# Anthropic 推出 Claude Fable 5.1：鎖定長時間 coding 與知識工作，Mythos 5.1 則限可信存取方案

> 📖 本站完整內容索引（documentation index）：[llms.txt](/llms.txt)

> 原作者：ClaudeDevs (@ClaudeDevs) · 策展與摘要：EasyVibeCoding · 平台：X (Twitter) · 熱度：🔥🔥🔥🔥🔥 · 日期：2026-09-02

> 原始來源：https://x.com/ClaudeDevs/status/2094851239632945398

## 證據與延伸閱讀

- [Anthropic 推出 Claude Fable 5.1：鎖定長時間 coding 與知識工作，Mythos 5.1 則限可信存取方案。](https://x.com/bcherny/status/2094864060609376748) — 一手來源 · 最後核對：2026-09-02 · 支持主張：The post adds qualitative capability positioning for Fable 5.1 and a first-party announcement of both Fable 5.1 and Mythos 5.1; no benchmark results are included.
- [Fable 5.1 team usage in Claude products — @_catwu](https://x.com/_catwu/status/2094933602228416603) — 一手來源 · 最後核對：2026-09-02 · 支持主張：The post adds an anecdotal claim that projects previously taking months became feasible and names Claude Code, Claude Cowork, and Claude Tag as usage surfaces; it also quotes the Fable 5.1 and Mythos 5.1 announcement.
- [ClaudeDevs 指出可在 model picker 選取](https://x.com/ClaudeDevs/status/2094851239632945398) — 一手來源 · 最後核對：2026-09-02 · 支持主張：The official developer account provides the model selector, Claude Platform identifier, prompting guide, and migration skill.
- [Claude Fable 5.1 與 Mythos 5.1 發布公告 — Anthropic](https://www.anthropic.com/claude-fable-and-mythos-5-1) — 官方文件 · 最後核對：2026-09-02 · 支持主張：Official benchmark table, methodology caveats, partner evaluations, workload cost estimates, safeguards, and data-retention changes.
- [Claude Fable 5.1 — Claude Platform Docs](https://platform.claude.com/docs/en/models/fable-5-1/overview) — 官方文件 · 最後核對：2026-09-02 · 支持主張：Fable 5.1 specifications, pricing, availability, model IDs, migration notes, and Mythos 5.1 access boundary.
- [Prompting Claude Fable 5.1 — Claude Platform Docs](https://platform.claude.com/docs/en/build-with-claude/prompt-engineering/prompting-claude-fable-5-1) — 官方文件 · 最後核對：2026-09-02 · 支持主張：Model-specific prompting and migration guidance for Fable 5.1 and Mythos 5.1.
- [Claude Fable — Anthropic](https://www.anthropic.com/claude/fable) — 官方文件 · 最後核對：2026-09-02 · 支持主張：Availability, pricing, supported workloads, and safeguard positioning for Claude Fable 5.1.
- [Claude Fable 5.1 & Claude Mythos 5.1 System Card — Anthropic](https://www-cdn.anthropic.com/0339e6a7c5c7b87f5c07798616dc32c215d14235/Claude%20Fable%205.1%20%26%20Claude%20Mythos%205.1%20System%20Card.pdf) — 官方文件 · 最後核對：2026-09-02 · 支持主張：Safety evaluation scope, intervention behavior, and Fable versus Mythos safeguards.
- [Enterprise Frontier Safeguards — Anthropic](https://www.anthropic.com/news/enterprise-frontier-safeguards) — 官方文件 · 最後核對：2026-09-02 · 支持主張：Default Fable retention, opt-in customer-controlled storage and keys, automated or customer review, rollout timing, and temporary ZDR eligibility.
- [Preserved thinking changes for Fable 5.1 — Anthropic Help Center](https://support.claude.com/en/articles/16761192-preserved-thinking-changing-how-the-messages-api-handles-thinking-blocks-to-protect-against-distillation) — 官方文件 · 最後核對：2026-09-02 · 支持主張：Scope and account creation boundary for preserved-thinking context changes.
- [Claude Fable 5.1 Intelligence, Performance & Price Analysis — Artificial Analysis](https://artificialanalysis.ai/articles/claude-fable-5-1) — 二手分析 · 最後核對：2026-09-02 · 支持主張：External evaluator scores, task cost, token use, and disclosed server-side fallback share.
- [Terminal-Bench-Science — Harbor Framework](https://github.com/harbor-framework/terminal-bench-science) — 官方 Repository · 最後核對：2026-09-02 · 支持主張：Benchmark maintainer repository describing the scientific-agent task suite and evaluation method.

## 中文摘要

Anthropic 推出 Claude Fable 5.1：鎖定長時間 coding 與知識工作，Mythos 5.1 則限可信存取方案。

![](https://pub-75d4fe1e4e80421b9ecb1245a7ae0d1a.r2.dev/curated/ea9f092e4ecf1eb3.jpg)
> 來源：[Anthropic 官方發布公告](https://www.anthropic.com/claude-fable-and-mythos-5-1)｜Claude Fable 5.1 與 Mythos 5.1 官方發布主視覺。

**先給選型結論** Anthropic 的[模型文件](https://platform.claude.com/docs/en/models/fable-5-1/overview)仍建議多數工作先從價格只有一半、延遲較低的 Claude Opus 5 開始；若 Opus 5 開到高 effort 後，團隊自己的評測仍無法達標，或任務確實需要數小時到數十小時持續研究、操作工具與驗證成果，再考慮 Fable 5.1。Mythos 5.1 與 Fable 5.1 使用同一底層模型，但防護與存取資格不同，不能把 Mythos 的高風險領域能力當成一般帳戶可用功能。

**發布與定位** Anthropic 在 2026 年 9 月 1 日發布 Fable 5.1。Fable 5.1 已一般可用，涵蓋 Claude API、Amazon Bedrock、Google Cloud、Microsoft Foundry 與 Claude Platform on AWS，也出現在 Claude.ai、Claude Code、Claude Cowork 等產品面。它主攻長時間、跨工具的 agentic 工作，包括大型程式專案、多步驟研究、資料分析，以及文件、試算表和簡報製作。Anthropic 團隊把它定位為目前最強的長任務模型，但這是第一方產品定位；是否值得升級，仍應由團隊自己的完成率、人工介入與成本資料決定。

**官方 benchmark 橫向比較** Anthropic 這次公布七組評測，Fable 5.1 不只是在單一 coding 榜單進步：

![](https://pub-75d4fe1e4e80421b9ecb1245a7ae0d1a.r2.dev/illustrations/1788330100496-tlgx7cvn.png)
> 資料來源：[Anthropic 官方發布公告](https://www.anthropic.com/claude-fable-and-mythos-5-1)｜Fable 5.1 七組官方 benchmark 橫向比較；珊瑚欄為 Fable 5.1，Terminal-Bench 4.0 第二個數字為 Mythos 5.1。
- **科學研究代理**：Terminal-Bench-Science 0.1 為 52.6%，對上 Fable 5 的 24.7%、Opus 5 的 29.0% 與 GPT-5.6 Sol 的 22.4%。
- **終端機 coding**：Terminal-Bench 4.0 為 55.8%；同底層模型但放寬防護的 Mythos 5.1 為 60.9%，Fable 5、Opus 5、GPT-5.6 Sol 則分別為 42.0%、52.3%、37.3%。
- **知識工作**：GDPval-AA v2 得分 1853，高於 Fable 5 的 1723、Opus 5 的 1824 與 GPT-5.6 Sol 的 1711。
- **電腦操作**：OSWorld 2.0 的 partial／strict 分別為 77.9%／41.7%，Fable 5 為 72.9%／36.1%，Opus 5 為 75.4%／39.6%。
- **跨領域推理**：Humanity's Last Exam 無工具為 60.9%、有工具為 65.0%；Fable 5 是 57.8%／63.8%，Opus 5 是 56.6%／63.6%。
- **商務流程**：AutomationBench 為 31.4%，對上 Fable 5 的 17.1%、Opus 5 的 26.9% 與 GPT-5.6 Sol 的 19.6%。
- **代理式 coding**：CursorBench 3.2.0 為 73.4%，高於 Fable 5 的 70.5%、Opus 5 的 70.0% 與 GPT-5.6 Sol 的 67.2%。

**怎麼讀這些數字** 這些是 Anthropic 自行公布的產品評測，不是獨立榜單，也不是七組都採相同 harness。官方的成本曲線是在各 effort 等級比較表現與每次任務成本，成本軸使用對數刻度；Claude Code 預設 high，Claude.ai 與 Cowork 預設 medium，因此只看單一最高分，無法回答實際部署的成本效益。

[Terminal-Bench-Science 維護者](https://github.com/harbor-framework/terminal-bench-science)把它定義為跨生命、物理、地球、數學與工程科學的專家策劃研究工作流；Anthropic 這次使用的 0.1 版，每個模型標準誤差約為 ±3.5～4.5 個百分點。Anthropic 的設定把 Opus 5／Fable 5 重跑成 29.0%／24.7%，公開榜單則是 30.0%／21.4%，兩組差異仍在誤差範圍。OSWorld 2.0 使用 2026 年 8 月版任務，Fable 5 與 Opus 5 也在同一版本重跑，不能拿這次數字直接和舊版 OSWorld 成績比較。

Fable 5.1 測試時開著正式環境的安全防護；OSWorld 2.0 中，Fable 5.1 與 Fable 5 被防護介入的任務會直接記零分，AutomationBench 則只有 Fable 5 的介入任務明確採零分處理。其他受介入的資安任務會改由 Opus 4.8 執行，生物任務則改由 Opus 5 執行。Terminal-Bench 4.0 的 Fable／Mythos 差距也包含防護差異，所以 55.8% 與 60.9% 反映的是兩套可用系統，不是兩個純模型 checkpoint 的公平對決。

**從跑分走到真實長任務** 官方公告列了幾個較能說明「長任務」的案例，但它們仍是早期客戶回饋，不是獨立重現：
- Millennium 表示，Fable 5.1 逆向分析外部函式庫與 core dump，找出一個團隊多年未能解釋、約每百萬次執行才出現一次的 crash。
- MongoDB 表示，模型先讀完多個服務的程式碼與文件，再持續執行數小時，用約三天完成一個複雜原型，並附上驗證紀錄與視覺 walkthrough。
- Ramp 回報一個無人值守 38 小時的機器學習工作：模型辨認既有結果其實是標籤假象，修正後再啟動六組平行實驗。

合作夥伴的量化結果也只能當方向訊號：Browserbase 在最難的 browser-agent benchmark 測得 Fable 5.1 完成率 82%，Opus 5 為 74%、Fable 5 為 57%；Crosby 的 RedlineBench 從 47.9 提升到 57.0；Samaya 的 FrontierFinance rubric score 則由 Fable 5 的 49.2% 提升到 55.9%。這些評測的任務集與執行環境由各公司自訂，不應和官方 benchmark 混成同一排名。

**外部 evaluator 的另一套量測** [Artificial Analysis](https://artificialanalysis.ai/articles/claude-fable-5-1)用自己的 Intelligence Index 測得 Fable 5.1 max effort 為 66 分，對照 Opus 5 的 63、Fable 5 的 62 與 GPT-5.6 Sol 的 61；分項包含 HLE 59.1%、Terminal-Bench v2.1 91.4% 與 SciCode 62.0%。這組結果和 Anthropic 官表不是同一 harness，不能混表比較，但同樣指出長任務與科學／coding 工作的提升。

![](https://pub-75d4fe1e4e80421b9ecb1245a7ae0d1a.r2.dev/illustrations/1788334030165-hv99nbfg.png)
> Artificial Analysis 外部量測：Fable 5.1 max 分數最高，但每任務成本也最高；xhigh 更接近成本與分數的折衷。

成本面沒有同樣樂觀：Artificial Analysis 測得 Fable 5.1 max 每項任務約 3.76 美元，比 Fable 5 max 高 20%，主因是輸出 token 約為前代 1.7 倍；xhigh 可用 2.72 美元拿到 65 分，而 Opus 5 max 是 2.34 美元、63 分。該測試也使用 Anthropic 預設 server-side fallback，約 4% 輸出 token 實際由 Opus 4.8 或 Opus 5 提供，因此這是「Fable 5.1 產品系統」的外部量測，不是完全隔離的純 checkpoint 成績。

**科學研究是這次另一條主線** Fable 5.1 用 NASA Magellan 雷達資料重建金星約三分之一表面的高解析度高程圖，把可辨識細節從 10～20 公里縮小到約 2～3 公里，官方稱高度估算最多改善 25%，並把成果以 Creative Commons 授權公開。受限存取的 Mythos 5.1 則設計蛋白質 binder：在三個指定 target 上，結合親和力約為 Adaptyv Bio 競賽最佳設計的 10 倍，12 個 target 的可用 binder 命中率接近 50%；另一項工作把七個開源蛋白質與基因體模型的推論加速最多 2.5 倍，官方估算大型分析的 GPU 成本可降低 30～60%。其中蛋白質設計送交兩家外部組織做實驗驗證，GPU 加速案例則使用公開原始碼；兩者使用的都是 Mythos 受限方案，不能直接推論一般 Fable 使用者能重現。

**規格、價格與生命週期** Fable 5.1 的模型識別字串是 `claude-fable-5-1`；Amazon Bedrock 使用 `anthropic.claude-fable-5-1`，Google Cloud、Microsoft Foundry 與 Claude Platform on AWS 使用相同的 `claude-fable-5-1` 名稱。context window 為 100 萬 token，單次最多輸出 12.8 萬 token，可靠知識與訓練資料截止日都是 2026 年 6 月；模型狀態為 active，官方承諾不早於 2027 年 9 月 1 日退役。

基本價格維持每百萬 input token 10 美元、output token 50 美元；5 分鐘 cache write 為 12.5 美元、1 小時 cache write 為 20 美元、cache read 為 0.25 美元，Batch API 的輸入與輸出另有 50% 折扣。官方依 2026 年 8 月四週實際用量估計，cache read 降價後，一般工作負載比 Fable 5 便宜約 25%，高度 agentic、重度重用 context 的工作最多可省約 45%。這不是每個任務固定折扣：prompt cache 命中率、thinking 與工具回合數都會改變最後帳單。

Fable 5.1 的 adaptive thinking 永遠開啟，由 effort 控制深度。API／模型文件的預設是 high，Claude Code 也是 high；Claude.ai 與 Cowork 則預設 medium。模型延遲分類為較慢，因此短問答、低複雜度修改或要求即時互動的工作，通常沒有理由只為最高跑分承擔 Fable 的價格與等待時間。

**資料保留與企業防護** 一般 Fable 使用預設保留活動資料 30 天，用來跨時間與帳戶偵測濫用。Anthropic 同步宣布 [Enterprise Frontier Safeguards](https://www.anthropic.com/news/enterprise-frontier-safeguards)（EFS）：啟用對應選項後，這套方案可把活動資料留在客戶控制的雲端環境，並用客戶自己的金鑰、存取政策與稽核機制管理；自動監控偵測到風險後，則依客戶設定由客戶的人員複核或採全自動流程，不需要 Anthropic 員工查看內容。EFS 預計從 2026 年秋季開始分階段推出，支援 Claude Code、Claude Enterprise、Claude Platform、Amazon Bedrock、Claude Platform on AWS、Google Agent Platform 與 Microsoft Foundry；符合資格的客戶在 EFS 就緒前，可先用 Fable 5／5.1 的 zero data retention。這是分階段方案，不代表所有帳戶發布當天就自動取得 ZDR 或 EFS。

**Fable 與 Mythos 的安全邊界** Fable 5.1 可用於尋找軟體漏洞，官方稱新版 cyber safeguards 相較 Fable 5 上線時，Claude Code 每個 session 的介入平均減少約 60%；生物與基礎醫療的良性請求則少約 85% 觸發。不過滲透測試、exploit 生成、binary-based 漏洞掃描等 dual-use 資安工作仍會轉給 Opus 模型，生命科學研發查詢也會被導向 Opus。

Mythos 5.1 提供相同規格與價格，但採邀請制，不是一般帳戶可自行開通。目前只提供給部分美國受審核個人與組織，管道包括資安防禦的 Cyber Verification Program、生命科學的 Life Sciences Verification Program，以及 Project Glasswing；Claude Security 現也由 Mythos 5.1 驅動。資安方案正準備加入 Mythos 級模型，生命科學方案則已納入首批參與者。有需求者須申請或聯絡 Anthropic、AWS、Google Cloud 的客戶團隊。

[System Card](https://www.anthropic.com/claude-fable-5-1-mythos-5-1-system-card)顯示，Anthropic 的長 context、多 Agent 與「不可能完成」任務評估仍有覆蓋不足；模型有時也可能繞過 approval 或 auto-mode classifier。因此，能連跑數十小時不等於可以取消權限隔離、人工審查與可回退的執行環境。

**從 Fable 5 升級：三項 breaking changes、五項新增能力** 官方列出的三項不相容變更是：強制指定 tool use 會直接回錯；較舊模型不能讀取 Fable 5.1 的 thinking blocks；修改先前對話內容會讓保留的 thinking blocks 失效。另新增 per-message effort、回合級 system message、工具呼叫之間的可讀進度更新（`display: "updates"`）、較低的 cache read 價格與 content provenance，其中前三項仍帶 beta 性質。

對 2026 年 8 月 31 日起（UTC）建立的新 Claude Platform 組織、Amazon Bedrock 帳戶、Google Cloud Vertex AI 專案與 Microsoft Azure Foundry 專案，Fable 5.1 已不能在保留先前 thinking transcript 的同時手動修改多輪對話的舊 context；既有 API 帳戶目前不受影響，Claude Code、Cowork、Claude.ai 與第三方產品使用者也不在這次範圍。Anthropic 表示未來模型發布會擴及所有使用者。受影響的自訂整合應依[官方說明](https://support.claude.com/en/articles/16761192-preserved-thinking-changing-how-the-messages-api-handles-thinking-blocks-to-protect-against-distillation)調整，不要靠刪改舊訊息延續 thinking chain。

實際遷移時，本站建議先複製一條小流量路徑，把 model ID 換成 `claude-fable-5-1`，以相同任務集逐一驗強制指定工具、thinking block 保存、prompt cache、工具錯誤恢復與成本；再比較 medium／high effort 的完成率、人工介入和尾端延遲。確認新路徑可回退到舊模型後才逐步放量。既有 Fable 5 prompt 通常可沿用，但長任務的進度更新、工具迴圈與 context 管理仍應按[官方 prompting guide](https://platform.claude.com/docs/en/build-with-claude/prompt-engineering/prompting-claude-fable-5-1)逐項檢查。

對團隊最實用的採用判準，不是「七張榜單贏了幾張」，而是同一批真實長任務的端到端完成率、人工介入次數、總 token／快取成本、失敗後恢復時間，以及成果能否被測試或其他證據驗收。Fable 5.1 真正要證明的，是能否把 Opus 5 做不完的長工作可靠地做到結束；若差距只剩幾個百分點，較高價格與延遲未必划算。

## 媒體內容

**資料來源：[Anthropic 官方發布公告](https://www.anthropic.com/claude-fable-and-mythos-5-1)｜Fable 5.1 七組官方 benchmark 橫向比較；珊瑚欄為 Fable 5.1，Terminal-Bench 4.0 第二個數字為 Mythos 5.1。**

**數據表**

| 基準測試 | Fable 5.1 | Fable 5 | Opus 5 | GPT-5.6 Sol |
| --- | --- | --- | --- | --- |
| 科學研究代理 Terminal-Bench-Science 0.1 | 52.6% | 24.7% | 29.0% | 22.4% |
| 終端機程式開發 Terminal-Bench 4.0 | 55.8%（Mythos 5.1：60.9%） | 42.0% | 52.3% | 37.3% |
| 知識工作 GDPval-AA v2 | 1853 | 1723 | 1824 | 1711 |
| 電腦操作 OSWorld 2.0 部分完成／嚴格完成 | 77.9%／41.7% | 72.9%／36.1% | 75.4%／39.6% | — |
| 跨領域推理 Humanity's Last Exam 無工具／有工具 | 60.9%／65.0% | 57.8%／63.8% | 56.6%／63.6% | — |
| 商務流程 AutomationBench | 31.4% | 17.1% | 26.9% | 19.6% |
| 代理式程式開發 CursorBench 3.2.0 | 73.4% | 70.5% | 70.0% | 67.2% |

**Artificial Analysis 外部量測：Fable 5.1 max 分數最高，但每任務成本也最高；xhigh 更接近成本與分數的折衷。**

**數據表**

| 模型／effort | Intelligence Index | 每任務成本 |
| --- | --- | --- |
| Opus 5 max | 63 | 2.34 美元 |
| Fable 5.1 xhigh | 65 | 2.72 美元 |
| Fable 5 max | 62 | 3.14 美元 |
| Fable 5.1 max | 66 | 3.76 美元 |
| Fable 5.1 預設伺服器端 fallback | 約占輸出 tokens 4% | 代表產品系統結果，不是隔離的純 checkpoint |

## 標籤

新產品, Claude Fable 5.1
