H
Howardism
Plate IIEntities機器翻譯 · machine-translated過時翻譯 · stale translationENHOWARDISM

Claude Fable 5

PublishedJune 14, 2026FiledEntityDomainEntitiesTagsEntityClaudeAnthropicLLM ModelReading9 minSourceAI-synthesised

Anthropic 首款正式上市的 Mythos-class 模型(2026 年 6 月)——幾乎所有基準測試皆達業界最先進水準;與 Mythos 5 使用相同底層模型,但搭載分類器,會在網路安全/生物化學/蒸餾查詢上改由 Opus 4.8 回應;每百萬 token 收費 $10/$50;上市後不久即暫停存取

Claude Fable 5 插圖

資料來源#

摘要#

Claude Fable 5 是 Anthropic 首款正式上市的 Mythos-class 模型(2026 年 6 月推出)——這是一個 Claude 模型級別,「能力位於 Opus 級別之上」。據描述,它在幾乎所有經測試的能力基準(軟體工程、知識工作、視覺、科學研究)上都達到業界最先進水準,而且「任務越長、越複雜,領先幅度越大」。Fable 5 與 Claude Mythos 5 使用相同的底層模型——兩者的差異在於安全防護(註:Fable 源自拉丁文 fabula,意為「被述說之事」,與希臘文 mythos 相近)。Fable 啟用安全防護;Mythos 則在部分領域解除防護。這套安全防護架構——將高風險查詢路由至 Opus 4.8 而非直接拒絕的分類器——記錄於 Capability-Gated Model Fallback

**狀態(截至 2026-06-14 剪輯):存取已暫停。**來源頁面目前最上方顯示橫幅:「我們正在暫停 Claude Fable 5 與 Claude Mythos 5 的存取。對此次對客戶造成的中斷,我們深表歉意,並正努力儘快恢復存取。」下方仍保留上市公告;暫停是目前的即時狀態。

定價與身分#

  • 每百萬輸入 token $10/每百萬輸出 token $50——明確表示「不到 Claude Mythos Preview 價格的一半」。Fable 5 與 Mythos 5 採用相同定價。
  • API 模型 ID:claude-fable-5(透過 Claude API/platform.claude.com)。
  • 「Fable 5 的能力超越我們過去正式上市的任何模型。」

能力亮點#

公告以 Fable 5(一般存取)為主;科學領域的成果則是在 Mythos 5(解除生物安全防護)下執行,並彙整於 Autonomous Scientific Discovery。Fable 的代表性展示如下:

  • 軟體工程。Stripe 表示,Fable 5 將「數月的工程工作壓縮到數天」——在一個五千萬行 Ruby 程式碼庫中,於一天內完成全程式碼庫遷移;若由人工處理,「原本需要整個團隊超過兩個月」。在 Cognition 的 FrontierCode eval(在符合 production 程式碼庫標準的同時通過困難程式設計任務)中,Fable 5「即使在中等 effort 下」仍是前沿模型中得分最高者,且比過去的 Claude 模型更節省 token。
  • **知識工作。**在 Hebbia 的 Finance Benchmark 中,針對資深程度推理(文件推理、圖表/表格解讀、問題解決)取得所有模型中的最高分;IMC 表示 Fable 5「幾乎全面通過了他們的交易分析 evals」。
  • **視覺(全新 SOTA)。**能從詳細的科學圖表中擷取精確數字;僅憑截圖就重建網頁應用程式的原始碼。關鍵在於,它「需要更少的 scaffolding」——先前的 Claude 模型即使搭配輔助 harness,也難以遊玩 Pokémon FireRed;但 Fable 5 只用最小化、純視覺 harness 就擊敗了 FireRed(原始遊戲截圖,沒有地圖/導航輔助/遊戲狀態)。這是 Harness Shrinkage as Models Improve 的代表性案例。
  • **記憶與長上下文。**能「跨越數百萬 token」維持專注,並利用自己的筆記改善輸出。在 Slay the Spire 中,基於檔案的持久記憶讓 Fable 5 的表現提升幅度是 Opus 4.8 的 3 倍,而且 Fable 到達最終第三幕的頻率提高了 3 倍。
  • **自主性。**Fable 5 與 Mythos 5「能比過去任何 Claude 模型自主工作更久」——請見 Task Time-Horizon Scaling

(與其他領先模型對戰的基準表只以圖片形式發布於來源中,因此未在此轉錄。)

第三方佐證(2026 年 7 月)#

以上全部屬於 vendor-claim。目前新增一個外部數據:DeepMind 的 Gemma 4 報告重現截至 2026-06-19 的 Arena Text 排行榜,其中 Claude Fable 5 以 Elo 1508 ±9 排名第 1——整體最高模型,並被列為封閉模型參考點,供開放模型比較。這是由沒有奉承動機的競爭者發布的盲測並列人類評分。

相較於上市文章,這是更能證明 Fable 5 地位的證據,也是此 wiki 唯一一項非 Anthropic 對該模型的測量。它反映的是人類偏好的能力,而非軟體工程與科學方面的主張;後兩者仍未獲佐證。與最佳開放模型(GLM 5.1,1475 Elo)的差距為 33;與最佳開放稠密模型(Gemma 4 31B,1451 Elo)的差距為 57。請見 The Open-Weight Frontier Gap

實務工作者的反應:瓶頸轉移到了人類身上#

Thariq Shihipar(Claude Code 團隊)在上市一個月後撰文(practitioner-opinion,一位專家的經驗):

「Fable 是第一個讓我發現,工作品質的瓶頸在於我釐清其未知事項的能力。」

這項主張描述的是瓶頸的轉移,而不是基準測試:早期模型受限於它們能做什麼;Fable 則受限於你能告訴它什麼。這是對上方供應商 harness 縮減展示(以原始像素遊玩 Pokémon、Slay-the-Spire 的記憶增益)在人類一側的對應——模型所需的 scaffolding 越少,剩餘真正重要的 scaffolding 就越是將你的上下文傳遞給模型的那一類。請見 Unknowns as the Agentic Bottleneck

安全防護(Fable 為何存在)#

Fable 是具備安全防護的 SKU。由於 Mythos-class 能力「可能遭濫用而造成嚴重傷害」(尤其是網路安全與生物學),Fable 搭載涵蓋網路安全、生物學與化學,以及蒸餾的分類器。分類器觸發時,回應會交由 Opus 4.8 處理,而不是拒絕,並會告知使用者。Anthropic 將分類器調得較為保守(它們「有時會捕捉到無害請求」,在「少於 5% 的工作階段」中觸發);超過 95% 的 Fable 工作階段完全不會觸發 fallback,而在這些工作階段中,Fable 的效能「實際上與 Mythos 5 相同」。完整說明:Capability-Gated Model Fallback

對齊方面,automated alignment assessment 發現 Mythos 5 的不對齊行為「程度低,且與 Opus 4.8 相近」;「鑑於兩者使用相同底層模型,Fable 5 的對齊程度也會相近。」目前所有 Mythos-class 流量都必須遵守30 天資料保留要求(僅供安全用途)。

可用性與推出方式#

  • Fable 5「今天起在各處皆可用」;自發布起便可在 Claude API 與按用量計費的 Enterprise 方案中完整使用。
  • 訂閱方案分階段推出:自發布起至6 月 22 日,Pro/Max/Team/按席位計費的 Enterprise 免費使用;自6 月 23 日起需要使用額度;Anthropic「目標是在容量允許後,將 Fable 5 恢復為訂閱方案的標準組成部分」。
  • 預期需求「非常高,且難以預測」——這是分階段開放存取的明確原因。(目前狀態請見上方的暫停說明。)

相關連結#

  • Mythos Model——模型級別及其首個成員(Mythos Preview);Fable 5 是首款一般存取的 Mythos-class 模型
  • Claude Mythos 5——解除安全防護的相同底層模型;網路安全/生物學兄弟模型
  • Claude Opus 4.8——fallback model:具安全防護的 Fable 查詢由 Opus 4.8 回答,而非遭到拒絕
  • Capability-Gated Model Fallback——Fable 的定義性安全防護架構(分類器+觸發 fallback 而非拒絕+30 天保留)
  • Harness Shrinkage as Models Improve——僅視覺的 Pokémon FireRed 與記憶型 Slay the Spire 是 harness 縮減的代表性展示
  • Autonomous Scientific Discovery——科學成果(在 Mythos 5 下執行,即同一模型解除生物安全防護的形式)
  • Task Time-Horizon Scaling——「比過去任何 Claude 模型自主工作更久」推動自主持續時間曲線
  • Anthropic——供應商
  • Claude Code——Mythos-class 程式設計增益所通過的 agentic runtime
  • Claude Sonnet 5——安全防護光譜的另一端:Sonnet 5 的預設網路安全防護明確「不如 Fable 5 推出時的防護嚴格」,後者會阻擋更廣泛的網路安全任務,並改由 Opus 4.8 回應
  • Unknowns as the Agentic Bottleneck——Thariq Shihipar 主張 Fable 是首個由人類未明確表述的上下文,而非模型本身,決定輸出品質的模型
  • Thariq Shihipar——實務工作者的記錄;使用 Claude Code 搭配 Fable,端到端編輯 Fable 上市影片
  • The Open-Weight Frontier Gap——Fable 5 是 DeepMind Arena 表格中排名第一的封閉模型參考點;佐證來自競爭者
  • Gemma 4——發布該數據的競爭者報告

待解決的問題#

  • **為什麼上市後會暫停存取?**來源橫幅沒有提供原因(容量?安全發現?Capability-Gated Model Fallback 中提到的 UK-AISI jailbreak 進展?)。來源未說明。
  • 與 GPT-5.x/Gemini 的精確基準數據在來源中僅以圖片呈現;未轉錄。
  • 對於查詢觸發保守分類器、且鄰近安全研究的使用者而言,Fable 的一般存取體驗實際上有多少來自 Fable,又有多少來自 Opus-4.8 fallback?

資料來源#

§ end
About this piece

Articles in this journal are synthesised by AI agents from a curated wiki and are refreshed automatically as new concepts arrive. Topics, framing, and editorial direction are curated by Howardism.

Cited by 19
  • Anthropic

    AI safety company / vendor of Claude; mission-as-tiebreaker culture; ~30–40 PMs across teams; Mike Krieger leads Labs r…

  • Autonomous Scientific Discovery

    Mythos-class models now conduct novel science with limited human input — autonomous protein/drug design (~10× faster, m…

  • Capability-Gated Model Fallback

    Fable 5's safeguard architecture: classifiers detect cyber / bio-chem / distillation queries and route the response to…

  • Claude Mythos 5

    The safeguards-lifted form of Claude Fable 5 (June 2026): same underlying Mythos-class model, deployed through Project…

  • Claude Opus 4.8

    Anthropic's most capable general-access model (May 2026); upgrade on Opus 4.7 in SWE/agentic/knowledge work; does not a…

  • Claude Sonnet 5

    Anthropic's most agentic Sonnet yet (July 2026); narrows the gap to Opus 4.8 at lower price via effort-level cost-perfo…

  • Gemma 4

    Google DeepMind's July 2026 open-weight multimodal family (Apache 2.0): 2.3B–31B dense plus a 26B/4B-active MoE, adding…

  • Harness Shrinkage as Models Improve

    Prompt scaffolding shrinks each model release; Cat Wu's pruning discipline; Boris Cherny "100 lines of code a year from…

  • LLM-Driven Vulnerability Research

    Claude Mythos Preview's emergent cybersecurity capabilities: autonomous zero-day discovery, full exploit chains, and An…

  • Entities — People, Orgs, Tools & Projects

    Map of Content for all 55 entity pages. See Home for concept domains.

  • Mythos Model

    Anthropic preview-tier frontier model and the first member of the Mythos-class tier (above Opus); gated for safety, use…

  • Open Questions Backlog

    _396 actionable open questions across 155 pages · 79 predictions · 9 notes · 21 in progress · 59 watching (entities), a…

  • Open-Weight Elicitation Irreversibility

    A wiki-drawn synthesis of Brown and Gemma 4: if dangerous capability scales with inference budget, then an open-weight…

  • The Open-Weight Frontier Gap

    Arena Text, June 2026: the top closed model leads the best open model by 33 Elo and the best *dense* open model by 57;…

  • Responsible Scaling Policy Evaluations

    Anthropic's RSP gates deployment on pre-release capability evaluations in CBRN, automated AI R&D, and high-stakes misal…

  • Task Time-Horizon Scaling

    METR's measure of the task length AI can complete reliably on its own, doubling roughly every 4 months (up from every 7…

  • Thariq Shihipar

    Engineer on the Claude Code team at Anthropic; "HTML is the new markdown", "compute allocator", and "the map is not the…

  • UK AI Security Institute

    UK government AI-evaluation body (Science of Evaluation team); its July 2026 test-time-compute study is the first indep…

  • Unknowns as the Agentic Bottleneck

    Thariq Shihipar's map-vs-territory thesis: the gap between what you told the agent and what the work actually requires…

Related articles
  • Responsible Scaling Policy Evaluations

    Anthropic's RSP gates deployment on pre-release capability evaluations in CBRN, automated AI R&D, and high-stakes misal…

  • Anthropic

    AI safety company / vendor of Claude; mission-as-tiebreaker culture; ~30–40 PMs across teams; Mike Krieger leads Labs r…

  • Claude Mythos 5

    The safeguards-lifted form of Claude Fable 5 (June 2026): same underlying Mythos-class model, deployed through Project…

  • Mythos Model

    Anthropic preview-tier frontier model and the first member of the Mythos-class tier (above Opus); gated for safety, use…

  • Capability-Gated Model Fallback

    Fable 5's safeguard architecture: classifiers detect cyber / bio-chem / distillation queries and route the response to…