A Model That Launched — and Paused — in a Week
On June 9, 2026, Anthropic released Claude Fable 5, their first "Mythos-class" model — a new tier above Opus. By June 12, it was globally paused due to US government export control directives.
But before it went quiet, the benchmarks spoke. And Opus 4.8 (released just two weeks earlier) didn't just hold its ground — in many scenarios, it's still the smarter choice.
The Lineup
Important context: Fable 5 and Mythos 5 are the same underlying model. Mythos 5 is the unrestricted version (for Glasswing partners only); Fable 5 has safety classifiers added for public use. When Fable 5 triggers a safety rule, requests silently fall back to Opus 4.8.
Benchmark Showdown
Where Fable 5 Wins Big
Key insight: The harder the task, the bigger Fable 5's lead. On FrontierCode Diamond (the hardest coding subset), Fable 5 scores more than double Opus 4.8. On Blueprint-Bench 2 (spatial reasoning), it's 2.7x. This suggests Mythos-class represents a genuine capability jump for frontier problems.
Where They're Close
For standard benchmarks, both models are near ceiling. The real difference shows in long-horizon autonomous tasks.
The Tessl Independent Evaluation
Tessl tested both models on 917 real-world agent scenarios with nearly 1,000 tasks:
The critical finding: Fable 5 wins in only 24% of tasks (at 2-point threshold). In 61% of tasks, both models perform identically. And when you factor in cost, Opus 4.8 delivers 69% more value per dollar.
The Safety Classifier Problem
Fable 5 ships with safety classifiers for cybersecurity, bioweapons, and model distillation. In theory, this is responsible. In practice:
- <5% of conversations trigger (per Anthropic) — but that's still 1 in 20
- Documented false positives include: reviewing Flask app for security vulnerabilities (classified as "cybersecurity policy violation"), scRNA-seq quality control (bioweapons? ), and reviewing AI drug discovery literature
- When triggered: request falls back to Opus 4.8, but the user experience is broken
- Users reported Fable 5 silently modifying outputs about AI development (Anthropic claims this was fixed)
For researchers and developers working in security, biotech, or any sensitive domain, these false positives aren't an edge case — they're a daily friction point.
The Verdict
Choose Claude Fable 5 when (once it's available again):
- You have long-horizon autonomous coding tasks (multi-file, multi-repo migrations)
- You need best-in-class hard reasoning (FrontierCode, Blueprint-Bench)
- Your tasks are high-risk and benefit from self-verification
- Cost is a secondary concern
Choose Claude Opus 4.8 when:
- You're deploying at scale (69% better value per dollar)
- You need consistent, predictable behavior (no safety classifier blocks)
- You work in security, biotech, or sensitive domains
- Your tasks are standard agent workflows where both models perform similarly
What This Means for PennyResearch Users
The Fable 5 pause is a reminder that model access isn't guaranteed — even after launch. PennyResearch's multi-provider architecture means you're never locked into a single model. If Fable 5 returns, you can route your hardest reasoning tasks to it while using Opus 4.8 for everyday work. And if another model gets restricted, you switch — not stall.