All corrections
X June 10, 2026 at 09:53 AM

x.com/Plinz/status/2064528835329286487

1 correction found

1
Claim
The public Mythos is the first model that is deliberately designed to perform worse on technical tasks than its predecessors.
Correction

Anthropic’s June 9 release was Claude Fable 5 for general use; Claude Mythos 5 remained restricted. Anthropic also says Fable 5 exceeds all of its previous public models on capability benchmarks, including software engineering, so it is not a model that generally performs worse than its predecessors on technical tasks.

Full reasoning

Anthropic’s own June 9 announcement contradicts this in two ways.

  1. The generally available model was not “public Mythos.” Anthropic said it was launching Claude Fable 5 for general use, while Claude Mythos 5 would be available only to a small group of cyberdefenders and infrastructure providers through Project Glasswing. So taken literally, there was no general-public release of “Mythos” itself.

  2. Anthropic says the public release is stronger, not weaker, than prior public models. In the same announcement, Anthropic says Fable 5’s capabilities “exceed those of any model we’ve ever made generally available,” that it is “state-of-the-art on nearly all tested benchmarks,” and that it shows “exceptional performance in software engineering.” Those statements directly conflict with the idea that the public release was designed to perform worse than its predecessors on technical tasks in general.

  3. Anthropic had already described an earlier public model as the first one with deliberately reduced cyber capabilities. In the April 16 announcement for Claude Opus 4.7, Anthropic wrote that Opus 4.7 was “the first such model” and said that during training it had “experimented with efforts to differentially reduce” its cyber capabilities. That means even if the post is referring specifically to deliberately reduced cyber-related capabilities, Anthropic’s own description says the first such public model was Opus 4.7, not the June 9 Mythos-class release.

What is true is that Anthropic added safeguards to Fable 5: some cyber/bio/distillation queries are routed to Opus 4.8 instead. But that is different from saying the model is the first one deliberately designed to perform worse than its predecessors on technical tasks overall.

2 sources
  • Claude Fable 5 and Claude Mythos 5 | Anthropic

    Today we're launching Claude Fable 5: a Mythos-class model that we've made safe for general use... Fable 5's capabilities exceed those of any model we've ever made generally available. It is state-of-the-art on nearly all tested benchmarks of AI capability, showing exceptional performance in software engineering... For a small group of cyberdefenders and infrastructure providers, we're also launching Claude Mythos 5... Mythos 5 will initially be deployed through Project Glasswing.

  • Introducing Claude Opus 4.7 | Anthropic

    Opus 4.7 is the first such model: its cyber capabilities are not as advanced as those of Mythos Preview (indeed, during its training we experimented with efforts to differentially reduce these capabilities)... although it is less broadly capable than our most powerful model, Claude Mythos Preview—it shows better results than Opus 4.6 across a range of benchmarks.

Model: OPENAI_GPT_5 Prompt: v1.16.0