Research report · June 10, 2026
⏱ Covers everything released June 9–10, 2026

Mythos & Fable 5:
the complete research report.

Mythos & Fable 5 — The Complete Research Report | June 2026

Everything Anthropic released, said, priced, restricted and implied — in one place. No hype, no frameworks. Verified numbers, direct quotes, and 20 sources you can check yourself.

Claude Mythos — the new model class
TL;DR — the whole story in 10 lines
  • 🆕 June 9, 2026: Anthropic released Claude Fable 5 + announced Claude Mythos 5 — a brand-new fourth model class above Opus.
  • 🧠 Fable 5 and Mythos 5 are the same model. Fable has safety guardrails; Mythos doesn't and is restricted to vetted partners.
  • 📊 It's state-of-the-art on nearly every benchmark — 95.0% SWE-bench Verified, more than double Opus 4.8 on frontier code.
  • 💰 $10 in / $50 out per million tokens — exactly 2× Opus 4.8, less than half the old Mythos Preview price.
  • 🎁 Included free in all paid Claude plans until June 22. From June 23 it needs usage credits.
  • 🛡️ Three topics auto-fall back to Opus 4.8: offensive cyber, bio/chem, model distillation — under 5% of sessions.
  • 🏛️ Mythos 5 goes to Project Glasswing: ~150 government-linked cyber-defence orgs in 15+ countries.
  • ⚠️ Days before launch, Anthropic publicly proposed a "coordinated brake pedal" for frontier AI.
  • 🤖 Anthropic disclosed that over 80% of its own merged code is now written by Claude.
  • 🔑 API model ID: claude-fable-5 · 1M token context · 128K output · live on the API, Claude Code and Amazon Bedrock.
01 · What was announced

Two models. One brain. A new class.

On June 9, 2026, Anthropic announced Claude Fable 5 and Claude Mythos 5 together.

This is the headline most coverage missed: Claude's family now has four classes — Haiku, Sonnet, Opus, and Mythos, which sits above Opus.

a.Claude Fable 5 — the public one

The first Mythos-class model the public can use. Anthropic's words: "a Mythos-class model that has been made safe for general use."

Available immediately via API (claude-fable-5), in Claude apps, Claude Code, and Amazon Bedrock.

b.Claude Mythos 5 — the restricted twin

The identical underlying model with the safeguards lifted.

Only available to approved users: Project Glasswing cybersecurity partners and, soon, vetted biology researchers.

Anyone on the old Mythos Preview gets auto-upgraded to Mythos 5.

▌The one-line takeaway

The distinction is purely the safeguards. Buy Fable 5 and you're running the same brain that governments get — with rails.

02 · How we got here

The timeline.

APRIL 2026

Mythos Preview launches quietly — restricted to a handful of partners over cybersecurity concerns. Priced at $25/$125 per MTok.

MAY 2026

Anthropic discloses that over 80% of its own merged code is authored by Claude; engineers ship ~8× more code per quarter than 2024.

JUNE 2, 2026

Mythos Preview access expands to hundreds of organizations across 15 countries managing critical infrastructure.

EARLY JUNE 2026

Anthropic publishes "When AI Builds Itself" — warning about recursive self-improvement and proposing a verifiable, coordinated slowdown mechanism among frontier labs.

JUNE 9, 2026

Claude Fable 5 released publicly. Mythos 5 announced for trusted access. Free on paid plans until June 22.

JUNE 23, 2026

The free window closes — Fable 5 usage moves to usage credits on subscription plans.

03 · The benchmarks

Every verified number, one place.

Software engineering

SWE-bench Verified
Fable 595.0%
Opus 4.888.6%
SWE-bench Pro
Fable 580.0%
Opus 4.869.2%
GPT-5.558.6%
FrontierCode Diamond (Cognition)
Fable 529.3
Opus 4.813.4

Plus: 72.9% on CursorBench at max effort — 8.6 points ahead of GPT-5.5.

Analytics, documents & vision

📈 First model ever past 90% on Hex's core analytics benchmark — a 10-point jump over Opus.

📄 OfficeQA Pro: 57.9% vs Opus 4.8's 48.1%.

👁️ GDP.pdf (hard document vision): 29.8% vs Opus 4.8 (22.5%), GPT-5.5 (24.9%), Gemini 3.1 Pro (16.7%). Anthropic calls it the new state of the art for vision.

Long context & agentic work

🧵 1M token context window, 128K max output.

🧵 GraphWalks BFS at 1M depth: 79.4 F1 vs Opus 4.8's 68.1.

🌐 BrowseComp multi-agent: 93.3%.

🛠️ Toolathlon: 61.7% in 19.8 average turns vs Opus 4.8's 59.9% in 24.5 turns — better results in fewer steps.

🔬 Matches or beats GPT-5.5 on frontier physics research using one-third the tokens.

💼 Real-World Finance v2: preferred in 74% of comparisons (Elo 1,374 vs 1,222) + top score on Hebbia's finance benchmark.

▌The pattern across every benchmark

The analysts' consensus line: "the lead widens as tasks get longer and more complex." Single questions barely change. Real work changes a lot.

04 · The wild capabilities

The stuff that sounds fake but is sourced.

🎮Plays Pokémon FireRed by raw vision alone — older models needed helper harnesses
~10×drug-design acceleration by Anthropic's internal experts
9 / 14protein targets yielded real drug design candidates
80%of the time, biologists preferred its novel hypotheses in blind tests
1 weekof autonomous novel genomics research, unsupervised
50Mlines — the Stripe codebase migration it worked across

"Compressed months of engineering into days."

— Stripe's CEO, on the 50-million-line migration, June 2026

05 · Pricing & access

What it costs. Exactly.

Opus 4.8
$5 / $25
Fable 5
$10 / $50
Mythos Preview (was)
$25 / $125

per million tokens, input / output

The access facts

🔑 API model ID: claude-fable-5 — live everywhere from day one.

📦 Batch pricing halves it: $5 in / $25 out — Opus money for Mythos-class output.

🎁 Included in Pro, Max, Team and Enterprise plans June 9–22 at no extra cost. From June 23: usage credits, pending capacity.

🎛️ New effort parameter: low, medium, high, and a new xhigh tier for the hardest jobs.

☁️ Available on Amazon Bedrock at launch.

🔒 All Mythos-class traffic: 30-day retention, never used for training, human access logged.

▌The practical read

Sticker price is 2× Opus — but it finishes jobs in ~20% fewer turns, and batch mode erases the difference entirely. For agent workloads, the finished-job cost is often a wash or better.

06 · The safeguards

How they made Mythos shippable.

Three classifier categories trigger a fallback

1. Cybersecurity — offensive hacking, vulnerability exploitation, defense evasion.

2. Biology & chemistry — bioweapons-adjacent and dual-use research.

3. Distillation — attempts to extract the model's capabilities.

When triggered, the request is answered by Claude Opus 4.8 instead — and you're told it happened. Developers see stop_reason: "refusal" with a category, and can configure server-side fallback.

<5%of sessions ever trigger the fallback
1,000+hours of external red-teaming — no universal jailbreak found
0 / 30harmful compliances across 30 public jailbreak techniques

The honest caveats — from Anthropic itself

⚠️ The safeguards are "deliberately tuned to be cautious" and "stricter than would be ideal" — benign requests will sometimes trip them.

⚠️ The UK AI Safety Institute made "initial progress" toward a jailbreak within a brief testing window.

⚠️ Anthropic concedes it's "likely impossible to completely prevent universal jailbreaks."

⚠️ The alignment assessment found Mythos 5's misalignment levels "similar to that of Opus 4.8" — better tools, same fundamental challenges.

07 · Mythos 5 & Project Glasswing

The version you can't have.

Mythos 5 — sealed behind trusted access

Who gets the unrestricted model

🏛️ Project Glasswing — Anthropic's collaboration with the US government: cyber-defenders and critical-infrastructure providers get Mythos 5 with the cyber safeguards lifted, immediately.

🌍 The program is expanding to roughly 150 organizations in more than fifteen countries.

🧬 Biomedical researchers get a trusted-access program "in the coming weeks" with bio/chem safeguards removed.

📈 A broader trusted-access program is promised later.

08 · The strange backstory

They asked for a brake pedal. Then they shipped.

Days before this release, Anthropic published an essay — "When AI Builds Itself" — that reads strangely next to a launch.

What the essay actually says

It maps three futures: capability gains plateau; AI keeps accelerating with humans steering; or AI achieves recursive self-improvement — "fully autonomously designing and developing its own successor."

That last one, they write, "is not inevitable" but "could come sooner than most institutions are prepared for."

The proposal: verification systems so frontier labs can prove to each other they've slowed down.

"If such systems existed, we expect that we would slow down or temporarily pause, if other developers at or near the frontier also did so in a verifiable manner."

— Anthropic, "When AI Builds Itself", June 2026

The stats buried in it

🤖 As of May 2026, over 80% of Anthropic's merged code is authored by Claude.

📈 Anthropic engineers ship roughly 8× more code per quarter than in 2024.

Their framing of why: "The doing... now costs almost nothing in human time, even if it still has costs in compute."

Read together: the company most loudly warning about acceleration is also the one publishing proof of its own — and shipping the strongest public model anyway, with the safeguards as its answer to the contradiction.

09 · Reactions & quotes

What people are actually saying.

"Fable 5's capabilities exceed those of any model we've ever made generally available."

— Anthropic, official announcement, June 9, 2026

"Highest-scoring model on FrontierBench... excels at long-horizon reasoning."

— Scott Wu, CEO of Cognition, June 2026

"A future where developers can hand increasingly ambitious work to agents and trust the results."

— Mario Rodriguez, Chief Product Officer, GitHub, June 2026

The press framing split into two camps: the capability story (VentureBeat: "brings Mythos to the masses... most powerful generally available model ever") and the tension story (TechCrunch: released "days after warning AI is getting too dangerous"; NBC: "built on the same tech that spooked the government").

10 · What it means for you

The down-to-earth read.

If you use Claude for work

Switch your hardest tasks to Fable 5 before June 22 while it's free on your plan. Test it on the jobs the old model fumbled — that's where the gap shows.

If you run agents or automations

Point one workflow at claude-fable-5 and watch turn counts drop. Keep Opus 4.8 as your fallback — that's literally Anthropic's own architecture inside Fable 5.

Use the effort dial: low for routine runs, xhigh for the jobs that used to fail.

If you're watching costs

Batch mode is the cheat code — $5/$25 is Opus pricing for Mythos-class output. For overnight or queued work, there's no reason to pay the live rate.

If you're thinking about the bigger picture

The 80%-of-code stat is the real headline. The frontier lab is already mostly AI-written. The question isn't whether agentic work arrives — it's whether your systems are ready to use it.

11 · Open questions

What we still don't know.

Capacity. Anthropic expects high demand and staged the rollout — how constrained will access be after June 23?

False-positive rates over time. The classifiers are "deliberately cautious" — will they loosen as data accrues?

The trusted-access expansion. Who counts as "trusted" for bio research, and how fast does it widen?

Competitive response. GPT-5.5 just got jumped on the hardest benchmarks — OpenAI and Google's counters are presumably imminent.

Subscription economics. "Usage credits" after June 22 is vague — real-world costs for heavy users are TBD.

12 · Sources

All 20 sources. Check everything.