資料來源#
- A Field Guide to Fable: Finding Your Unknowns
- Claude Fable 5 and Claude Mythos 5
- Gemma 4 Technical Report
摘要#
Claude Fable 5 是 Anthropic 首款正式上市的 Mythos-class 模型(2026 年 6 月推出)——這是一個 Claude 模型級別,「能力位於 Opus 級別之上」。據描述,它在幾乎所有經測試的能力基準(軟體工程、知識工作、視覺、科學研究)上都達到業界最先進水準,而且「任務越長、越複雜,領先幅度越大」。Fable 5 與 Claude Mythos 5 使用相同的底層模型——兩者的差異僅在於安全防護(註:Fable 源自拉丁文 fabula,意為「被述說之事」,與希臘文 mythos 相近)。Fable 啟用安全防護;Mythos 則在部分領域解除防護。這套安全防護架構——將高風險查詢路由至 Opus 4.8 而非直接拒絕的分類器——記錄於 Capability-Gated Model Fallback。
**狀態(截至 2026-06-14 剪輯):存取已暫停。**來源頁面目前最上方顯示橫幅:「我們正在暫停 Claude Fable 5 與 Claude Mythos 5 的存取。對此次對客戶造成的中斷,我們深表歉意,並正努力儘快恢復存取。」下方仍保留上市公告;暫停是目前的即時狀態。
定價與身分#
- 每百萬輸入 token $10/每百萬輸出 token $50——明確表示「不到 Claude Mythos Preview 價格的一半」。Fable 5 與 Mythos 5 採用相同定價。
- API 模型 ID:
claude-fable-5(透過 Claude API/platform.claude.com)。 - 「Fable 5 的能力超越我們過去正式上市的任何模型。」
能力亮點#
公告以 Fable 5(一般存取)為主;科學領域的成果則是在 Mythos 5(解除生物安全防護)下執行,並彙整於 Autonomous Scientific Discovery。Fable 的代表性展示如下:
- 軟體工程。Stripe 表示,Fable 5 將「數月的工程工作壓縮到數天」——在一個五千萬行 Ruby 程式碼庫中,於一天內完成全程式碼庫遷移;若由人工處理,「原本需要整個團隊超過兩個月」。在 Cognition 的 FrontierCode eval(在符合 production 程式碼庫標準的同時通過困難程式設計任務)中,Fable 5「即使在中等 effort 下」仍是前沿模型中得分最高者,且比過去的 Claude 模型更節省 token。
- **知識工作。**在 Hebbia 的 Finance Benchmark 中,針對資深程度推理(文件推理、圖表/表格解讀、問題解決)取得所有模型中的最高分;IMC 表示 Fable 5「幾乎全面通過了他們的交易分析 evals」。
- **視覺(全新 SOTA)。**能從詳細的科學圖表中擷取精確數字;僅憑截圖就重建網頁應用程式的原始碼。關鍵在於,它「需要更少的 scaffolding」——先前的 Claude 模型即使搭配輔助 harness,也難以遊玩 Pokémon FireRed;但 Fable 5 只用最小化、純視覺 harness 就擊敗了 FireRed(原始遊戲截圖,沒有地圖/導航輔助/遊戲狀態)。這是 Harness Shrinkage as Models Improve 的代表性案例。
- **記憶與長上下文。**能「跨越數百萬 token」維持專注,並利用自己的筆記改善輸出。在 Slay the Spire 中,基於檔案的持久記憶讓 Fable 5 的表現提升幅度是 Opus 4.8 的 3 倍,而且 Fable 到達最終第三幕的頻率提高了 3 倍。
- **自主性。**Fable 5 與 Mythos 5「能比過去任何 Claude 模型自主工作更久」——請見 Task Time-Horizon Scaling。
(與其他領先模型對戰的基準表只以圖片形式發布於來源中,因此未在此轉錄。)
第三方佐證(2026 年 7 月)#
以上全部屬於 vendor-claim。目前新增一個外部數據:DeepMind 的 Gemma 4 報告重現截至 2026-06-19 的 Arena Text 排行榜,其中 Claude Fable 5 以 Elo 1508 ±9 排名第 1——整體最高模型,並被列為封閉模型參考點,供開放模型比較。這是由沒有奉承動機的競爭者發布的盲測並列人類評分。
相較於上市文章,這是更能證明 Fable 5 地位的證據,也是此 wiki 唯一一項非 Anthropic 對該模型的測量。它反映的是人類偏好的能力,而非軟體工程與科學方面的主張;後兩者仍未獲佐證。與最佳開放模型(GLM 5.1,1475 Elo)的差距為 33;與最佳開放稠密模型(Gemma 4 31B,1451 Elo)的差距為 57。請見 The Open-Weight Frontier Gap。
實務工作者的反應:瓶頸轉移到了人類身上#
Thariq Shihipar(Claude Code 團隊)在上市一個月後撰文(practitioner-opinion,一位專家的經驗):
「Fable 是第一個讓我發現,工作品質的瓶頸在於我釐清其未知事項的能力。」
這項主張描述的是瓶頸的轉移,而不是基準測試:早期模型受限於它們能做什麼;Fable 則受限於你能告訴它什麼。這是對上方供應商 harness 縮減展示(以原始像素遊玩 Pokémon、Slay-the-Spire 的記憶增益)在人類一側的對應——模型所需的 scaffolding 越少,剩餘真正重要的 scaffolding 就越是將你的上下文傳遞給模型的那一類。請見 Unknowns as the Agentic Bottleneck。
安全防護(Fable 為何存在)#
Fable 是具備安全防護的 SKU。由於 Mythos-class 能力「可能遭濫用而造成嚴重傷害」(尤其是網路安全與生物學),Fable 搭載涵蓋網路安全、生物學與化學,以及蒸餾的分類器。分類器觸發時,回應會交由 Opus 4.8 處理,而不是拒絕,並會告知使用者。Anthropic 將分類器調得較為保守(它們「有時會捕捉到無害請求」,在「少於 5% 的工作階段」中觸發);超過 95% 的 Fable 工作階段完全不會觸發 fallback,而在這些工作階段中,Fable 的效能「實際上與 Mythos 5 相同」。完整說明:Capability-Gated Model Fallback。
對齊方面,automated alignment assessment 發現 Mythos 5 的不對齊行為「程度低,且與 Opus 4.8 相近」;「鑑於兩者使用相同底層模型,Fable 5 的對齊程度也會相近。」目前所有 Mythos-class 流量都必須遵守30 天資料保留要求(僅供安全用途)。
可用性與推出方式#
- Fable 5「今天起在各處皆可用」;自發布起便可在 Claude API 與按用量計費的 Enterprise 方案中完整使用。
- 訂閱方案分階段推出:自發布起至6 月 22 日,Pro/Max/Team/按席位計費的 Enterprise 免費使用;自6 月 23 日起需要使用額度;Anthropic「目標是在容量允許後,將 Fable 5 恢復為訂閱方案的標準組成部分」。
- 預期需求「非常高,且難以預測」——這是分階段開放存取的明確原因。(目前狀態請見上方的暫停說明。)
相關連結#
- Mythos Model——模型級別及其首個成員(Mythos Preview);Fable 5 是首款一般存取的 Mythos-class 模型
- Claude Mythos 5——解除安全防護的相同底層模型;網路安全/生物學兄弟模型
- Claude Opus 4.8——fallback model:具安全防護的 Fable 查詢由 Opus 4.8 回答,而非遭到拒絕
- Capability-Gated Model Fallback——Fable 的定義性安全防護架構(分類器+觸發 fallback 而非拒絕+30 天保留)
- Harness Shrinkage as Models Improve——僅視覺的 Pokémon FireRed 與記憶型 Slay the Spire 是 harness 縮減的代表性展示
- Autonomous Scientific Discovery——科學成果(在 Mythos 5 下執行,即同一模型解除生物安全防護的形式)
- Task Time-Horizon Scaling——「比過去任何 Claude 模型自主工作更久」推動自主持續時間曲線
- Anthropic——供應商
- Claude Code——Mythos-class 程式設計增益所通過的 agentic runtime
- Claude Sonnet 5——安全防護光譜的另一端:Sonnet 5 的預設網路安全防護明確「不如 Fable 5 推出時的防護嚴格」,後者會阻擋更廣泛的網路安全任務,並改由 Opus 4.8 回應
- Unknowns as the Agentic Bottleneck——Thariq Shihipar 主張 Fable 是首個由人類未明確表述的上下文,而非模型本身,決定輸出品質的模型
- Thariq Shihipar——實務工作者的記錄;使用 Claude Code 搭配 Fable,端到端編輯 Fable 上市影片
- The Open-Weight Frontier Gap——Fable 5 是 DeepMind Arena 表格中排名第一的封閉模型參考點;佐證來自競爭者
- Gemma 4——發布該數據的競爭者報告
待解決的問題#
- **為什麼上市後會暫停存取?**來源橫幅沒有提供原因(容量?安全發現?Capability-Gated Model Fallback 中提到的 UK-AISI jailbreak 進展?)。來源未說明。
- 與 GPT-5.x/Gemini 的精確基準數據在來源中僅以圖片呈現;未轉錄。
- 對於查詢觸發保守分類器、且鄰近安全研究的使用者而言,Fable 的一般存取體驗實際上有多少來自 Fable,又有多少來自 Opus-4.8 fallback?
資料來源#
- Claude Fable 5 and Claude Mythos 5——Anthropic,「Claude Fable 5 and Claude Mythos 5」(2026 年 6 月;AAV 編輯於 2026 年 6 月 9 日)
- A Field Guide to Fable: Finding Your Unknowns——Thariq Shihipar,2026-07-04(
practitioner-opinion):瓶頸轉移主張與上市影片建置記錄 - Gemma 4 Technical Report——表 4(
empirical):第三方 Arena Text 表現,截至 2026-06-19 以 Elo 1508 ±9 排名第 1
Cited by 19
- Anthropic
AI safety company / vendor of Claude; mission-as-tiebreaker culture; ~30–40 PMs across teams; Mike Krieger leads Labs r…
- Autonomous Scientific Discovery
Mythos-class models now conduct novel science with limited human input — autonomous protein/drug design (~10× faster, m…
- Capability-Gated Model Fallback
Fable 5's safeguard architecture: classifiers detect cyber / bio-chem / distillation queries and route the response to…
- Claude Mythos 5
The safeguards-lifted form of Claude Fable 5 (June 2026): same underlying Mythos-class model, deployed through Project…
- Claude Opus 4.8
Anthropic's most capable general-access model (May 2026); upgrade on Opus 4.7 in SWE/agentic/knowledge work; does not a…
- Claude Sonnet 5
Anthropic's most agentic Sonnet yet (July 2026); narrows the gap to Opus 4.8 at lower price via effort-level cost-perfo…
- Gemma 4
Google DeepMind's July 2026 open-weight multimodal family (Apache 2.0): 2.3B–31B dense plus a 26B/4B-active MoE, adding…
- Harness Shrinkage as Models Improve
Prompt scaffolding shrinks each model release; Cat Wu's pruning discipline; Boris Cherny "100 lines of code a year from…
- LLM-Driven Vulnerability Research
Claude Mythos Preview's emergent cybersecurity capabilities: autonomous zero-day discovery, full exploit chains, and An…
- Entities — People, Orgs, Tools & Projects
Map of Content for all 55 entity pages. See Home for concept domains.
- Mythos Model
Anthropic preview-tier frontier model and the first member of the Mythos-class tier (above Opus); gated for safety, use…
- Open Questions Backlog
_396 actionable open questions across 155 pages · 79 predictions · 9 notes · 21 in progress · 59 watching (entities), a…
- Open-Weight Elicitation Irreversibility
A wiki-drawn synthesis of Brown and Gemma 4: if dangerous capability scales with inference budget, then an open-weight…
- The Open-Weight Frontier Gap
Arena Text, June 2026: the top closed model leads the best open model by 33 Elo and the best *dense* open model by 57;…
- Responsible Scaling Policy Evaluations
Anthropic's RSP gates deployment on pre-release capability evaluations in CBRN, automated AI R&D, and high-stakes misal…
- Task Time-Horizon Scaling
METR's measure of the task length AI can complete reliably on its own, doubling roughly every 4 months (up from every 7…
- Thariq Shihipar
Engineer on the Claude Code team at Anthropic; "HTML is the new markdown", "compute allocator", and "the map is not the…
- UK AI Security Institute
UK government AI-evaluation body (Science of Evaluation team); its July 2026 test-time-compute study is the first indep…
- Unknowns as the Agentic Bottleneck
Thariq Shihipar's map-vs-territory thesis: the gap between what you told the agent and what the work actually requires…
Related articles
- Responsible Scaling Policy Evaluations
Anthropic's RSP gates deployment on pre-release capability evaluations in CBRN, automated AI R&D, and high-stakes misal…
- Anthropic
AI safety company / vendor of Claude; mission-as-tiebreaker culture; ~30–40 PMs across teams; Mike Krieger leads Labs r…
- Claude Mythos 5
The safeguards-lifted form of Claude Fable 5 (June 2026): same underlying Mythos-class model, deployed through Project…
- Mythos Model
Anthropic preview-tier frontier model and the first member of the Mythos-class tier (above Opus); gated for safety, use…
- Capability-Gated Model Fallback
Fable 5's safeguard architecture: classifiers detect cyber / bio-chem / distillation queries and route the response to…
