Claude Mythos 5.1: what it is, who gets access, and what to run

Alicia Kirana Utomo
Written by

Alicia Kirana Utomo

Katelin Teen
Reviewed by

Katelin Teen

Last edited September 1, 2026

Expert Verified
An illustration of a vault door being opened by a small approved list of researchers, representing invite-only access to Claude Mythos 5.1

Can you actually get Claude Mythos 5.1?

Start here, because for most readers this is the whole question.

Claude Mythos 5.1 access check

Pick the description closest to your team. Every answer is drawn from Anthropic's own access pages as of September 2026.

Who are you?

Maybe, and not yet today

The Cyber Verification Program is the route. It already grants reduced cyber safeguards on Opus and Sonnet-class models, and Anthropic says Mythos-class access joins it in the near future. Applications go through the CVP portal.

Run today: Fable 5.1. It now identifies software vulnerabilities in source code, and Claude Code users see roughly 60% fewer cyber-safeguard interventions per session than on Fable 5.

Only by invitation

The Life Sciences Verification Program is an invite-only beta, built with the US government, that relaxes biology safeguards for professional R&D. First participants are enrolled; broader enrolment is promised but unscheduled.

Run today: Fable 5.1, whose biology safeguards now fire 85% less often on benign elementary biology and medical questions. Research-grade life sciences queries still route to Opus.

Ask your account team

Claude Mythos 5.1 is listed on the Claude API, Amazon Bedrock, Google Cloud and Microsoft Foundry, and the docs say to request access through your Anthropic, AWS, or Google Cloud account team. Availability is still restricted to a set of US organizations.

Run today: Fable 5.1 on the same platform, with Enterprise Frontier Safeguards arriving in phases this fall and zero data retention available to eligible customers in the meantime.

No, and that is unlikely to change soon

There is no self-serve path to Mythos 5.1. Both programs are organizational, vetted, and US-scoped. Nothing about a credit card or a usage tier moves you into them.

Run today: Fable 5.1 for the ceiling, or Opus 5 at $5 in and $25 out, which is what Anthropic's own docs tell most workloads to start on.

No, and you do not want it

Mythos 5.1's extra headroom is in cybersecurity and biology. Neither shows up in a refund request or a WISMO ticket. It is also the slowest and most expensive model in the family, on a queue where latency is what the customer feels.

Run today: Sonnet 5 or Opus 5 behind a layer that actually knows your product. The gap on a support queue is grounding and testing, not raw model tier.

What Claude Mythos 5.1 actually is

Claude Mythos 5.1 is Anthropic's Mythos-class frontier model, released on September 1, 2026 with the model ID claude-mythos-5-1. Its status in the docs reads "Active (invite only)", and retirement is committed no sooner than September 1, 2027.

The Claude Mythos product page on Anthropic's site listing the September 1, 2026 Mythos 5.1 announcement, as taken from Anthropic
The Claude Mythos product page on Anthropic's site listing the September 1, 2026 Mythos 5.1 announcement, as taken from Anthropic

The Mythos line started as Claude Mythos Preview, the unreleased model that Anthropic used to launch Project Glasswing in April 2026. Mythos 5 followed on June 9, went unavailable on June 12, and had its export controls lifted on July 1 after US government approval. Mythos 5.1 is the third entry in that line, and the first that arrives alongside a generally available twin.

Here is the spec sheet, taken from the model page rather than the marketing copy.

SpecClaude Mythos 5.1
Model IDclaude-mythos-5-1
ReleasedSeptember 1, 2026
StatusActive, invite only
Context window1M tokens
Max output128K tokens
ThinkingAdaptive, always on
Default efforthigh
Knowledge cutoffJune 2026
Comparative latencySlower
Input / outputText and images, to text
Input price$10 / MTok
Output price$50 / MTok
Cache read$0.25 / MTok
PlatformsClaude API, Bedrock, Google Cloud, Microsoft Foundry
RetirementNot sooner than September 1, 2027
The Claude Mythos 5.1 model page in Anthropic's platform docs showing pricing and specifications, as taken from Claude Platform Docs
The Claude Mythos 5.1 model page in Anthropic's platform docs showing pricing and specifications, as taken from Claude Platform Docs

One more line worth reading twice: using Mythos 5.1 requires accepting a 30-day data retention policy for safety monitoring by default. That is a real procurement fact, and it is the mirror image of the zero-retention story Anthropic is telling on the Fable side.

Mythos 5.1 and Fable 5.1 are one model with two locks

This is the piece that confuses almost everyone, including the Hacker News thread on launch day:

Hacker News

"Claude Mythos 5.1 is identical to Fable 5.1, but it offers more permissive safeguards for vetted individuals and organizations"

Then why does it have separate datapoints for Terminal Bench, and score higher? Something doesn't add up here??

It does add up, and the answer is mechanical rather than mysterious. Anthropic evaluates Fable 5.1 with production safeguards enabled. When those safeguards fire on a benchmark task, the task does not simply fail; it gets handed to a different model. Cybersecurity tasks were completed by Opus 4.8, biology tasks by Opus 5, and on OSWorld 2.0 the intervened tasks scored a flat zero. So the published Fable 5.1 number is the number for the whole safeguarded system, not for the weights.

Mythos 5.1 is those same weights with the cyber and biology classifiers loosened. That is the entire delta.

A diagram showing one September 2026 model splitting into Fable 5.1, generally available with safeguards on, and Mythos 5.1, invite only with safeguards relaxed
A diagram showing one September 2026 model splitting into Fable 5.1, generally available with safeguards on, and Mythos 5.1, invite only with safeguards relaxed

Anthropic is explicit that it expects the gap to shrink. The launch post notes that the Terminal-Bench difference reflects the earlier, less precise cyber safeguards, and that with the Fable 5.1 improvements shipping the same day, the difference between the two models should be much smaller.

What actually changed on the Fable side is worth knowing even if you never touch Mythos:

  • Fable 5.1 can now be used to identify software vulnerabilities in source code, which the previous safeguards blocked.
  • Claude Code users should see around 60% fewer cyber-safeguard interventions per session.
  • Biology safeguards fire 85% less often on benign elementary biology and medical questions than the ones that launched with Fable 5.
  • Dual-use work is still redirected to Opus models: penetration testing, exploit generation, and binary-based vulnerability scanning.

That last bullet is the honest boundary. Fable 5.1 will read your source and tell you where the bug is. It will not write you the exploit.

The benchmarks, and the one that matters

Anthropic published Fable 5.1 and Mythos 5.1 numbers side by side with Fable 5, Opus 5, and GPT-5.6 Sol. Scores below are Fable 5.1 unless the row says otherwise.

BenchmarkFable 5.1Fable 5Opus 5GPT-5.6 Sol
Terminal-Bench-Science 0.152.6%24.7%29.0%22.4%
Terminal-Bench 4.055.8% (60.9% Mythos 5.1)42.0%52.3%37.3%
GDPval-AA v21853172318241711
OSWorld 2.0 (strict)41.7%36.1%39.6%not published
Humanity's Last Exam (tools)65.0%63.8%63.6%not published
AutomationBench31.4%17.1%26.9%19.6%
CursorBench 3.2.073.4%70.5%70.0%67.2%

The standout is Terminal-Bench-Science, which more than doubled against Fable 5. Everything else is a narrower step, and the launch thread noticed:

Hacker News

The price reduction comes from the cache read pricing falling from $1/M to $0.25/M, which means that Fable 5.1 now costs half of Opus's cache read costs ($0.5/M).

That reading is fair on the coding rows and less fair on the science row, where a 27.9-point jump is not a rounding artefact. Anthropic footnotes a standard error of 3.5 to 4.5 points per model on that benchmark, which still leaves the gap comfortably outside noise.

The science demos behind it are the part I would actually screenshot. Mythos 5.1 designed protein binders that beat the best entries in Adaptyv Bio's competitions by 10x on binding affinity for three targets, with a hit rate near 50% across 12 targets where 10 to 15% is typical. It also wrote GPU kernels that sped up seven open-source genomics and protein models by up to 2.5x with identical outputs, cutting estimated GPU cost on genome-wide analyses by 30 to 60%. Separately, Fable 5.1 built a new high-resolution elevation map for a third of Venus from 30-year-old Magellan radar data.

None of that is a support use case. All of it is why the model is behind a door.

Pricing: the cache read cut is the actual news

Mythos 5.1 and Fable 5.1 carry the same rate card. Input is $10 per million tokens, output is $50 per million. That is unchanged from Fable 5.

What changed is the cache read line, and it changed a lot.

MeterFable 5 / Mythos 5Fable 5.1 / Mythos 5.1
Base input$10 / MTok$10 / MTok
Output$50 / MTok$50 / MTok
5-minute cache write$12.50 / MTok$12.50 / MTok
1-hour cache write$20 / MTok$20 / MTok
Cache read$1.00 / MTok$0.25 / MTok
Batch input / output$5 / $25$5 / $25

Every other Claude model prices a cache hit at 0.1x base input. Fable 5.1 and Mythos 5.1 price it at 0.025x, and Anthropic's docs call that out in a footnote as the only exception in the family. It is also, as the thread above noted, half of what an Opus 5 cache read costs, on a model that is twice the base price.

A bar chart comparing indexed cost on Fable 5 versus Fable 5.1, showing 100 to 75 on a typical workload and 100 to 55 on a highly agentic workload, with a note that cache reads fell from $1.00 to $0.25
A bar chart comparing indexed cost on Fable 5 versus Fable 5.1, showing 100 to 75 on a typical workload and 100 to 55 on a highly agentic workload, with a note that cache reads fell from $1.00 to $0.25

Anthropic measured this over four weeks of real August 2026 usage at default effort. Typical workloads across Claude Enterprise, Claude Code, and the API land around 25% cheaper. Context-heavy, tool-heavy agentic work, where cache reads are most of the bill, lands around 45% cheaper.

The practical consequence is a planning change, not just a smaller invoice:

Hacker News

"Cache reads now cost 75% less, or $0.25 per million tokens." For me, at a typical 95% cache hit rate, I think my optimal context window size before autocompaction goes from ~200K to ~400K tokens. Great for longer horizon tasks.

That is the whole point of the change. On a long agentic run, the transcript gets re-read on every turn, so the cache read meter is the one that compounds. If you have been aggressively compacting context to keep a bill down, the math for that just moved. The wider family rate card lives in my Anthropic API pricing breakdown, and the subscription side is in Claude pricing.

One thing the price cut does not change: long context is still billed flat. Claude 4.6 and later models include the full 1M window at standard rates, so a 900k-token request bills at the same per-token rate as a 9k-token one.

How access actually works

Two programs, both vetted, both US-scoped today.

Cyber Verification Program. The CVP currently provides reduced cyber safeguards on certain Opus and Sonnet-class models for defensive security work. Anthropic says Mythos-class access joins it in the near future, which is a promise rather than a date. Applications go through the Anthropic programs portal.

Life Sciences Verification Program. The LSVP is an invite-only beta, developed with the US government, that relaxes biology safeguards for professional research and development while leaving every other safeguard in place. First participants are enrolled, and Anthropic says it plans to expand to the broader life science community.

Existing enterprise customers have a third door: the docs tell you to contact your Anthropic, AWS, or Google Cloud account team. And Claude Security, Anthropic's codebase scanning product, now runs on Mythos 5.1, which is the one way to consume the model without being invited to it directly.

The Project Glasswing page describing the initiative and its launch partners, as taken from Anthropic
The Project Glasswing page describing the initiative and its launch partners, as taken from Anthropic

The reason for all of this sits in the Glasswing announcement from April 2026. Mythos Preview found thousands of zero-day vulnerabilities across every major operating system and browser, including a 27-year-old remote crash in OpenBSD and a 16-year-old flaw in FFmpeg that automated testing had hit five million times without catching. Anthropic committed $100M in usage credits and $4M in donations, and brought in twelve launch partners plus 40 more infrastructure organizations, expanding to roughly 150 organizations in over fifteen countries by June.

On the safety side, Anthropic's own assessment is that Mythos 5.1 has the strongest cyber capabilities of any model it has released, still sits in the lower risk category of its Frontier Compliance Framework, and falls short of the next risk tier in the Responsible Scaling Policy. It is deployed with the same biology safeguards as Mythos 5.

Three changes that will break someone's integration

Worth flagging even if Mythos itself is out of reach, because these land on Fable 5.1 too.

  1. Forced tool use returns an error. If your code sets tool_choice to any or a named tool, that call now fails. This is the one most likely to page someone.
  2. Thinking blocks are tied to the model that produced them. Earlier models cannot read Fable 5.1 or Mythos 5.1 thinking blocks, and editing an earlier turn invalidates them.
  3. Anti-distillation changes multi-turn editing. New API accounts created from launch day onward can no longer manually edit Claude's prior context while preserving the transcript of its prior thinking. Existing accounts are unaffected for now, but the change applies to everyone on future model releases.

There is also a compliance change with a long tail. To meet the EU AI Act's Code of Practice on Transparency of AI-Generated Content, models released after August 2, 2026 carry a text watermark, with a detection API in private preview for regulators, researchers, and obligated enterprises. Anthropic says it is invisible without the detection API and carries no information about the user or their conversations. It drew the most heat of anything in the launch thread, which is worth knowing if your content pipeline has an opinion about provenance. My notes on AI content and E-E-A-T cover the publishing side of that.

What to run instead, by job

Nearly everyone reading this is choosing between models they can actually call. Here is how I would split it.

If the job is customer conversations rather than code or research, the tier question is close to irrelevant. My roundup of the best model for support tickets walks the actual tradeoff, and custom AI models covers when training your own is worth it.

What a frontier model does not fix on a support queue

I want to be careful here, because the Mythos 5.1 results are real and the science is hard to argue with. But there is a failure pattern I have watched enough times to name it.

A team reads a launch post, upgrades the model behind their support bot, and expects resolution rate to move. It does not. What moves resolution rate is whether the model can see the right knowledge, whether it knows when to stop and hand off, and whether anyone tested it against real tickets before a customer saw it. None of those are model-tier problems.

The trust half of this is the one that costs money. As one CX lead at a DTC supplements brand put it on a sales call:

"The AI will never be able to answer 100% of the questions... I need an AI who is only handling the tickets that it's confident to handle and all the other ones, leave them alone."

That is a scoping and control problem, and a 60.9% Terminal-Bench score has nothing to say about it. Neither does the build-your-own route, which is where a lot of teams land after a launch like this one. Karel at GENERAL BYTES put the tradeoff plainly:

"We could try to write our own LLM application but we didn't want to invest our time into that. We wanted something that we would not have to maintain."

The API key is the cheap part. The retrieval, the permission boundaries, the escalation rules, the evaluation harness, and the maintenance of all four as your product changes are the expensive part. That work does not get cheaper because cache reads did.

Where eesel fits

eesel sells AI teammates, not a model. The AI helpdesk teammate joins the queue you already run, reads the help centre and past tickets you already have, and answers within scope you set. The AI blog writer does the same job for content. Both arrive with the integrations and company context for their role, which is exactly the layer a raw frontier model does not ship with.

The eesel AI dashboard showing helpdesk activity across connected channels
The eesel AI dashboard showing helpdesk activity across connected channels

The specific thing I would point at, given this post's subject: before an eesel teammate answers a single live customer, it runs against your own historical tickets so you can see what it would have said. That is the dry run nobody gets from a model launch, and it is the difference between shipping an upgrade and shipping a guess. It connects to Zendesk, Freshdesk, Gorgias, Front, Help Scout, and HubSpot in minutes, and it is free to try.

Chasing Mythos 5.1 for a support queue is optimising the one variable that was never the constraint. Connecting Claude to your helpdesk is a more useful afternoon, and using Claude for support is the honest version of what that gets you.

My verdict on Claude Mythos 5.1

It is the most capable model Anthropic has shipped, it is not available to you, and the version that is available to you is the same model. That is an unusual thing to be able to write, and it is the single most useful fact in this post.

If you are a cyberdefender or a life scientist, apply, because the science results justify the paperwork. If you are anyone else, the launch that matters is Fable 5.1's cache read price, which quietly made long-context agentic work about 45% cheaper. That will change more production systems this quarter than a locked model ever will.

Frequently Asked Questions

What is Claude Mythos 5.1?
Claude Mythos 5.1 is Anthropic's most capable frontier model, released on September 1, 2026 with the model ID claude-mythos-5-1. It is the same underlying model as Fable 5.1, shipped with looser cybersecurity and biology safeguards, and it is offered by invitation only through Project Glasswing. The earlier Claude Mythos explainer covers how the family started.
How much does Claude Mythos 5.1 cost?
Claude Mythos 5.1 pricing is $10 per million input tokens and $50 per million output, matching Fable 5.1 exactly. A cache read is $0.25 per million, a five-minute cache write is $12.50, a one-hour cache write is $20, and the Batch API halves input and output to $5 and $25. The rest of the family sits in my Anthropic API pricing breakdown.
How do I get access to Claude Mythos 5.1?
There are two routes, and both are gated. Cyberdefenders apply to the Cyber Verification Program, which Anthropic says will include Mythos-class models in the near future. Life scientists go through the Life Sciences Verification Program, currently an invite-only beta run with the US government. Existing enterprise customers can also ask their Anthropic, AWS, or Google Cloud account team. Access is limited to a set of US organizations today.
What is the difference between Claude Mythos 5.1 and Claude Fable 5.1?
The weights are identical. The difference is the safeguard layer wrapped around them: Fable 5.1 routes dual-use biology and some cybersecurity work to Opus models, and Mythos 5.1 does not. That gap shows up on benchmarks where the safeguards fire, which is why Mythos 5.1 scores 60.9% on Terminal-Bench 4.0 against 55.8% for Fable 5.1. My Opus 5 versus Fable 5 comparison explains the tier logic underneath both.
Is Claude Mythos 5.1 available on Bedrock or Google Cloud?
Yes, for invited organizations. The docs list Claude Mythos 5.1 on the Claude API, Amazon Bedrock, Google Cloud, and Microsoft Foundry, with the same claude-mythos-5-1 ID on each. Being on the platform is not the same as being able to call it, since the invitation is still required.
What should I use instead of Claude Mythos 5.1?
For nearly every team, Fable 5.1 is the same model without the waiting list, and Claude Opus 5 is half the token price and Anthropic's own default recommendation. If you are picking a model for tickets rather than for research, my roundup of the best model for support tickets is the more useful comparison.
Does Claude Mythos 5.1 have a 1M context window?
Yes. Claude Mythos 5.1 carries a 1M token context window with 128K max output, adaptive thinking that is always on, and a default effort level of high. Anthropic bills the whole window at standard rates, so a 900k-token request costs the same per token as a 9k-token one. Claude Code context sizing covers what that means day to day.
Is Claude Mythos 5.1 safe to point at customer data?
Using Claude Mythos 5.1 requires accepting a 30-day data retention policy for safety monitoring by default. Anthropic's Enterprise Frontier Safeguards, which keep data in customer-controlled cloud infrastructure, roll out in phases starting this fall, and eligible customers can run Fable 5.1 with zero data retention until then. If the destination is a support queue, the model choice matters far less than the grounding and testing layer around it.

Share this article

Alicia Kirana Utomo

Article by

Alicia Kirana Utomo

Kira is a writer at eesel AI with a Computer Science background and over a year of hands-on experience evaluating AI-powered customer service tools. She focuses on breaking down how helpdesk platforms and AI agents actually work so that support teams can make better buying decisions.

Related Posts

All posts →
Editorial illustration for a guide to what Claude Fable 5 can do, Anthropic's most powerful AI model
Guides

What can Claude Fable 5 do? A capability-by-capability guide

What can Claude Fable 5 do? Run for days unattended, write and ship code, read 1M-token documents, and check its own work. Here's what that means in practice.

Riellvriany IndriawanRiellvriany IndriawanJun 17, 2026
Editorial illustration of Claude Opus 4.8, Anthropic's flagship AI model
Guides

What is Claude Opus 4.8? A clear-eyed look at Anthropic's flagship model

Claude Opus 4.8 is Anthropic's latest flagship model. Here's what changed, what it costs, and what a smarter model actually means for AI customer support.

Riellvriany IndriawanRiellvriany IndriawanJun 17, 2026
Image alt text
Guides

An overview of Claude Opus 4.6 pricing and capabilities

Explore our deep dive into Claude Opus 4.6 pricing. We break down the costs, new features, and practical use cases for Anthropic's latest AI model.

Katelin TeenKatelin TeenFeb 6, 2026
Illustration of a person weighing a small low-cost AI model against a larger caped flagship model on pedestals
Trending

Claude Opus 5 vs Fable 5: which should you actually run?

Fable 5 costs exactly double Opus 5. I went through both system cards, the docs and the independent benchmarks to work out when that second dollar buys anything.

Rama Adi NugrahaRama Adi NugrahaJul 27, 2026
An illustration of a person selecting one of five reasoning effort dials on a Claude Opus 5 style control strip
Trending

Claude Opus 5: what it is and how to actually run it

Claude Opus 5 is not one model, it is five. A plain guide to the specs, the effort dial that decides your bill, and the two API changes that break old code.

Kurnia Kharisma Agung SamiadjieKurnia Kharisma Agung SamiadjieAug 5, 2026
Claude Skills vs Subagent: What’s the difference?
Guides

Claude Skills vs Subagent: What’s the difference?

Explore the detailed breakdown of Claude Skills vs Subagent. We cover how they work, their best use cases, and why tools like eesel AI offer a more practical approach for non-technical teams to build specialized AI assistants.

Stevia PutriStevia PutriOct 16, 2025
Illustration of a developer at a laptop watching an agentic coding loop run through code, checks and a bot
Trending

Claude Opus 5 review: near-frontier coding at half the price

A hands-on Claude Opus 5 review: what the benchmarks actually say, the hallucination rate that went up, and whether it belongs on a live support queue.

Alicia Kirana UtomoAlicia Kirana UtomoJul 27, 2026
Illustration comparing a heavyweight reasoning model against a fast balanced model on cost and capability
Trending

Claude Opus 5 vs Sonnet 5: which one should you use?

Claude Opus 5 costs 1.7x Sonnet 5 per token and still finishes some jobs cheaper. Here is the head-to-head on price, benchmarks and real cost per task.

Rama Adi NugrahaRama Adi NugrahaJul 27, 2026
A practical guide to the new Claude create files feature
Guides

A practical guide to the new Claude create files feature

Anthropic’s Claude now creates files like Excel sheets and PowerPoints. Useful for quick tasks, but risky for business-critical automation. Here’s the full breakdown.

Kenneth PanganKenneth PanganSep 9, 2025

Ready to hire your AI teammate?

Set up in minutes. No credit card required.

Get started free