New scrutiny centers on the model’s opaque reasoning, adding transparency and interpretability concerns to earlier criticism over censorship and false positives.
Anthropic previously apologized for censorship issues involving Claude Fable 5 and said it would make fixes. Fresh criticism now focuses on the model’s opaque reasoning, with concerns that the complexity of its internal logic makes its outputs harder for users to interpret and trust. The combined debate links moderation problems, false positives and broader transparency questions that continue to shape expectations for how AI companies explain model behavior.