OpenAIAnthropic NewsJun 9, 2026, 12:00 AM

Claude Fable 5 and Claude Mythos 5

A condensed section focused on the key takeaways first.

Original Post

Quick Digest

Summary

A condensed section focused on the key takeaways first.

openaienmodel: gpt-5-mini-2025-08-07

Claude Fable 5 and Claude Mythos 5

Key Points

  • State-of-the-art capability across coding, vision, and long-context tasks
  • Conservative safeguards route risky queries to Claude Opus 4.8 (<5% sessions)
  • Mythos 5: lifted safeguards for vetted cyberdefense and accelerated protein design

Summary

Today’s launch introduces Claude Fable 5 (a Mythos-class model tuned for safe general use) and Claude Mythos 5 (the same underlying model with lifted safeguards for vetted cyberdefenders). Fable 5 sets a new state-of-the-art on benchmarks for software engineering, vision, knowledge work, long-context memory, and life-sciences research. To reduce misuse risk, Claude routes some sensitive queries to Claude Opus 4.8 via conservative safeguards (triggering in <5% of sessions). Mythos 5 is being deployed initially through Project Glasswing and a trusted-access program for infrastructure and cyberdefense users, and has powered ~10x speedups in internal protein design work and novel scientific hypotheses.

Key Points

  • Capability: Fable 5 exceeds prior generally available models across long-horizon tasks, coding, vision, and complex reasoning; Mythos 5 matches these capabilities with fewer safeguards in specific areas.
  • Safeguards & routing: Risky topics are conservatively blocked or answered by Opus 4.8; expect occasional false positives while safeguards are improved.
  • Access model: Fable 5 is broadly available; Mythos 5 is restricted to Project Glasswing participants and a phased trusted access program.
  • Practical guidance for engineers:
    • Choose Fable 5 for general-purpose engineering, agentic workflows, vision-based tasks, and knowledge work where high capability and safety are required.
    • Expect Opus 4.8 fallback for queries flagged as sensitive (e.g., offensive cybersecurity techniques); design tooling to detect and handle fallback responses.
    • Apply for or coordinate with authorized programs to access Mythos 5 for advanced cybersecurity, infrastructure defense, or sensitive life-sciences workflows.
  • Pricing: $10 per million input tokens and $50 per million output tokens (noted as < half the price of Claude Mythos Preview).
  • Known limits: Safeguards are conservative and may block benign requests; alignment testing reports low levels of misaligned behavior comparable to Opus 4.8.

Actionable notes

  • When integrating Fable 5 into CI/CD or agent pipelines, add logic to surface Opus 4.8 fallback responses and preserve audit logs for flagged sessions.
  • For sensitive security or bio workflows, engage with Project Glasswing / the trusted-access program and include strict access controls and monitoring.
  • Expect iterative improvements to safeguards; validate workflows against edge cases that might be falsely flagged.

Full Translation

Translations

A translation section that keeps the flow of the original article.

openaijamodel: gpt-5-mini-2025-08-07

Claude Fable 5 と Claude Mythos 5

発表 — Claude Fable 5 と Claude Mythos 5 (2026-06-09)

本日、一般利用向けに安全化した Mythos-class 1 モデルである Claude Fable 5 を公開します。Fable 5 は、これまで一般提供したどのモデルよりも優れた能力を持ちます。ソフトウェア工学、ナレッジワーク、ビジョン、科学研究など、ほぼすべてのテスト済みベンチマークで最先端の性能を示しています。タスクが長く複雑になるほど、Fable 5 の優位性は当社の他モデルに比べて大きくなります。

非常に高性能なモデルを公開することにはリスクが伴います。安全対策がなければ、サイバーセキュリティなどの分野で Fable 5 の能力が悪用されて深刻な被害を招く可能性があります。したがって、本モデルは一部のトピックに関する問い合わせに対しては、代わりに当社の次に高能力なモデルである Claude Opus 4.8 からの応答が返されるような安全策を導入して公開しています。安全策は迅速かつ安全に公開するために保守的に調整しており、無害なリクエストを誤検知する場合もありますが、セッション全体の平均で発動するのは 5% 未満です。今後数か月でさらに高能力なモデルが登場するにあたり、誤検知を減らすために安全策の改善を急いで行っています。

一部のサイバー防御担当者およびインフラ提供者向けに、保護領域を一部解除した同一の基盤モデルである Claude Mythos 5 もローンチします。Mythos 5 は初期は Claude Mythos Preview のアップグレードとして、米国政府との共同 Project Glasswing を通じて配備されます。世界のどのモデルよりも強力なサイバーセキュリティ能力を持ちます。近く、より広い信頼アクセスプログラムを通じたアクセス拡大を予定しています。

Fable 5 および Mythos 5 は、入力 100 万トークンあたり $10、出力 100 万トークンあたり $50 で提供されます — Claude Mythos Preview の価格の半分以下です。本日の共同ローンチは、高度な AI 能力をできるだけ多くのユーザーに、できるだけ速く、安全に届けるという目標に向けた一歩です。


Fable 5 / Mythos 5 の評価

以下は Fable 5 と Mythos 5 を他の先進モデルと比較した要約です。Fable 5 と Mythos 5 は、これまでの Claude 系モデルよりもより長時間にわたって自律的に動作できます。以下では、これらのスキルがソフトウェア工学にどう適用されるかを中心に、ナレッジワーク、ビジョン、メモリ、ライフサイエンス研究におけるモデルの向上点を述べます。

ソフトウェア工学

  • 初期テストでは、Stripe が Fable 5 によって数か月分のエンジニアリング作業が数日に圧縮されたと報告しています。5000万行の Ruby コードベースにおいて、通常はチーム全員で 2 か月以上かかるようなコードベース全体の移行を 1 日で実行しました。
  • Fable 5 は過去の Claude モデルよりトークン効率が高いです。Cognition の FrontierCode 評価(高品質なプロダクションコードベースの基準を満たしつつ困難なコーディング課題を解けるかをテスト)では、中程度の努力量でも Frontier モデルの中で最高スコアを記録しました。

ナレッジワーク

  • Fable 5 は複雑な分析タスクで高い性能を示します。Hebbia’s Finance Benchmark(上級レベルの推論を問う)では、Fable 5 が全モデル中で最高スコアを獲得し、文書ベースの推論、チャートや表の解釈、問題解決で大幅な向上を示しました。
  • IMC は Fable 5 がトレーディング分析の評価で、事実照会、概念的推論、根本原因解析、期待値解析など、ほぼ全域で優れた成績を収めたと報告しました。

ビジョン

  • Fable 5 はビジョンを含むタスクで新たな最先端モデルです。詳細な科学図表から正確な数値を抽出したり、スクリーンショットだけからウェブアプリのソースコードを再構築するなどの複雑なビジョンベースのタスクが可能です。
  • 以前の Claude モデルは追加の補助ツール(ハーネス)があっても Pokémon FireRed をうまくプレイできないことがありましたが、Fable 5 は最小限の「ビジョンのみ」のハーネスで FireRed をクリアしました。

メモリと長文コンテキスト

  • Fable 5 は長期間のタスクで数百万トークンにわたり集中を保ち、自身のメモを利用して出力を改善します。例として、デッキ構築ゲーム Slay the Spire をプレイさせた際、永続的なファイルベースのメモリにアクセスさせることで、Opus 4.8 に比べてパフォーマンスが 3 倍改善し、最終章に到達する頻度も 3 倍になりました。

実例(抜粋)

  • 天体シミュレーション: Claude Fable 5 は物理の第一原理から惑星の軌道運動を導出し、日食を予測するような太陽系シミュレーションを構築しました。
  • Factorio 自律プレイ: エンジニアに愛される工場構築ゲーム Factorio を自律的にプレイし、戦略を立てて自動化工場を構築しました。
  • 3D モデリング: ブラウザベースの CAD エディタ内で完全な 3D-プリント可能モデルを設計しました。エディタ自体(および組み込みの AI コパイロット)も Fable 5 によって作られました。
  • 音楽同期流体シミュレーション: Fable 5 は古典音楽の EDM リミックスに同期した流体シミュレーションをコードで作成しました(Fable 5 自身は音楽を事前に聞いていません)。

薬剤設計(Mythos 5)

  • Mythos 5 を用いた内部プロテインデザインの専門家チームは、薬剤設計プロセスの一部を約 10 倍高速化しました。
  • ある例では、Mythos 5 がプロテイン設計ツールとバイオインフォマティクスツールを用い、人間の介入なしで熟練オペレーターに匹敵または上回る成果を出しました。モデルは結合部位の選定、プロテイン設計ツールの選択と実行、途中での失敗からの回復など、通常は科学者が実行するタスクをすべて実行しました。
  • 研究で扱った 14 のタンパク標的のうち 9 件は強い候補を生成し、現在さらに調査中です(免疫チェックポイント、成長因子・受容体シグナル、神経変性、筋疾患、構造的に難しい標的など)。

分子生物学における新奇仮説

  • Mythos 5 は、一貫して新規で説得力のある科学仮説を生成する初のモデルです。ブラインドの一対一比較で、当社の科学者は約 80% の確率で Opus-class モデルより Mythos の分子生物学的仮説を好み、いくつかは実験評価へ進められています。
  • その間に、ある Mythos の仮説(E. coli タンパク質に関する新しい機構)が、同じ問題に独立に取り組んでいた研究室の研究で裏付けられました。

ゲノミクスにおける新規研究

  • Mythos 5 は主に自律的に 1 週間以上にわたり新しいゲノミクス研究を行いました。138 種の動物にまたがる数百万個の単一細胞データを組み上げ、遠縁の生物でも同じ役割を果たす細胞を識別するためのカスタム機械学習モデルを設計・訓練しました。
  • 最小限の高レベルの人間による入力だけで、Mythos 5 が訓練したモデルは、Science 誌に掲載された最近のモデルよりも高性能でした(モデルサイズは 100 倍小さいにもかかわらず)。これらの結果は数か月以内に論文化する予定です。

アライメント

  • 自動化されたアライメント評価では、Mythos 5 の不整合行動のレベル(欺瞞やユーザーによる悪用に協力するなどのモデルによる不適切な行動を含む)は低く、Opus 4.8 と同程度でした。Fable 5 と Mythos 5 は同一の基盤モデルであるため、Fable 5 のアライメントレベルも類似すると見なしています。
  • 評価の詳細およびその他の安全性・能力テストの詳細はモデルの system card に記載されています(全体の不整合行動レベルの図表は system card のセクション 6.2.3.1 を参照)。

早期アクセス顧客からのフィードバック(抜粋)

以下は、早期アクセスを受けた顧客が自社で行ったテストからの抜粋です。

  • "Claude Fable 5 is the state of the art model on CursorBench. It's opened up a class of long-horizon problems that were out of reach for earlier models." — Michael Truell, CEO and Co-founder
  • "Claude Fable 5 is a real step forward for the developers GitHub serves... a future where developers can hand increasingly ambitious work to agents and trust the results across the software lifecycle." — Mario Rodriguez, Chief Product Officer
  • "These are the strongest results of any Claude model we've had the opportunity to test. Claude Fable 5 is a clear step forward on agentic coding and prototyping." — Matt Colyer, Director of Product, Developers
  • "Claude Fable 5's reasoning is a clear step beyond Opus 4.8. It works at senior research scientist grade..." — Sean Ward, CEO and Co-founder
  • "Claude Fable 5 understands what builders mean, not just what they type. Apps that took a hundred prompts a year ago, it now one-shots." — Fabian Hedin, CTO & Co-founder
  • "Claude Fable 5 feels materially different. In blind review, our lawyers found its redlines matched or beat our current model every time." — Aveek Duttagupta, Member of Technical Staff
  • "At the highest effort, Claude Fable 5 reflects on and validates its own work... the extra thinking pays for itself." — Yusuke Kaji, GM, AI for Business
  • "Claude Fable 5 delivers more capable engineering in fewer turns than prior models..." — Luke Anderson, CTO
  • "Claude Fable 5 is the highest-scoring model on FrontierBench... It excels at long-horizon reasoning and generalizes to unfamiliar tools out of the box." — Scott Wu, CEO
  • "Claude Fable 5 is the strongest finance-first model we've tested..." — Damian Miraglia, Principal Engineer, Applied AI
  • "Claude Fable 5 is the first to break 90% on our core analytics benchmark of complex, long-running analytical tasks..." — Izzy Miller, AI Research Lead
  • "Claude Fable 5 is the strongest model we've tested on frontier physics research while using a third of the reasoning tokens..." — Matthew Pines, CEO
  • "On ViBench, Claude Fable 5 is the highest-performing model we've tested..." — Michele Catasta, President & Head of AI
  • "Claude Fable 5 beats Opus 4.8 on our everyday spreadsheet suite at every effort level... finishing runs 25–30% faster." — Peter Wang, Chief Science Officer

Claude Fable 5 の新しい安全策

Mythos-class モデルは重大なリスクを引き起こす閾値に達しています。今年 4 月に Project Glasswing を開始し、初の Mythos-class モデル(Claude Mythos Preview)を、サイバー防御担当者と重要なソフトウェアインフラ提供者の限定グループに提供しました。公開時に我々は、我々が hop