anthropic

Claude Fable 5.1

Claude Fable 5.1 是 Anthropic 用于 agentic coding 和科学计算的 flagship 模型,具有 1M context window、adaptive thinking 和 128K 输出 tokens。

AnthropicAgentic CodingLong ContextReasoningMultimodal
anthropic logoanthropicClaude2026年9月1日
上下文
1Mtokens
最大输出
128Ktokens
输入价格
$10.00/ 1M
输出价格
$50.00/ 1M
模态:TextImage
能力:视觉工具流式传输推理
基准测试
GPQA
93.7%
GPQA: 研究生级科学问答. 由领域专家创建的448道多选题的严格基准测试,涵盖生物学、物理学和化学。博士专家仅达到65-74%的准确率。 Claude Fable 5.1 在此基准测试中得分 93.7%。
HLE
60.9%
HLE: 高级专业推理. 测试模型在专业领域展示专家级推理能力的能力。 Claude Fable 5.1 在此基准测试中得分 60.9%。
MMLU
89.8%
MMLU: 大规模多任务语言理解. 涵盖57个学科的16,000道多选题的综合基准测试。 Claude Fable 5.1 在此基准测试中得分 89.8%。
MMLU Pro
78.4%
MMLU Pro: MMLU专业版. MMLU的增强版本,包含12,032道使用更难的10选项多选格式的问题。 Claude Fable 5.1 在此基准测试中得分 78.4%。
SimpleQA
44.5%
SimpleQA: 事实准确性基准. 测试模型对直接问题提供准确、事实性回答的能力。 Claude Fable 5.1 在此基准测试中得分 44.5%。
IFEval
88.6%
IFEval: 指令遵循评估. 衡量模型遵循特定指令和约束的能力。 Claude Fable 5.1 在此基准测试中得分 88.6%。
AIME 2025
88%
AIME 2025: 美国数学邀请赛. 来自著名AIME考试的竞赛级数学问题。 Claude Fable 5.1 在此基准测试中得分 88%。
MATH
94.2%
MATH: 数学问题解决. 涵盖代数、几何、微积分等领域的综合数学基准测试。 Claude Fable 5.1 在此基准测试中得分 94.2%。
GSM8k
97.6%
GSM8k: 小学数学8K. 8,500道需要多步推理的小学水平数学应用题。 Claude Fable 5.1 在此基准测试中得分 97.6%。
MGSM
92.1%
MGSM: 多语言小学数学. GSM8k基准测试翻译成10种语言版本。 Claude Fable 5.1 在此基准测试中得分 92.1%。
MathVista
71.5%
MathVista: 数学视觉推理. 测试解决涉及图表、图形等视觉元素的数学问题的能力。 Claude Fable 5.1 在此基准测试中得分 71.5%。
SWE-Bench
58.2%
SWE-Bench: 软件工程基准. AI模型尝试解决开源Python项目中的真实GitHub问题。 Claude Fable 5.1 在此基准测试中得分 58.2%。
HumanEval
92.4%
HumanEval: Python编程问题. 164道手写编程问题,模型必须生成正确的Python函数实现。 Claude Fable 5.1 在此基准测试中得分 92.4%。
LiveCodeBench
74.8%
LiveCodeBench: 实时编程基准. 在持续更新的真实世界编程挑战中测试编程能力。 Claude Fable 5.1 在此基准测试中得分 74.8%。
MMMU
71.2%
MMMU: 多模态理解. 大规模多学科多模态理解基准测试,测试视觉语言模型在大学水平问题上的表现。 Claude Fable 5.1 在此基准测试中得分 71.2%。
MMMU Pro
58.7%
MMMU Pro: MMMU专业版. MMMU的增强版本,问题更具挑战性,评估更严格。 Claude Fable 5.1 在此基准测试中得分 58.7%。
ChartQA
89.2%
ChartQA: 图表问答. 测试理解和推理图表信息的能力。 Claude Fable 5.1 在此基准测试中得分 89.2%。
DocVQA
94.1%
DocVQA: 文档视觉问答. 测试从文档图像中提取信息的能力。 Claude Fable 5.1 在此基准测试中得分 94.1%。
Terminal-Bench
55.8%
Terminal-Bench: 终端/CLI任务. 测试执行命令行操作和编写shell脚本的能力。 Claude Fable 5.1 在此基准测试中得分 55.8%。
ARC-AGI
12.4%
ARC-AGI: 抽象与推理. AGI抽象和推理语料库 - 通过新颖的模式识别谜题测试流体智力。 Claude Fable 5.1 在此基准测试中得分 12.4%。

关于 Claude Fable 5.1

了解 Claude Fable 5.1 的功能、特性以及它如何帮助您获得更好的效果。

架构概览

Claude Fable 5.1 是 Anthropic 构建的 flagship frontier 模型,专为长周期 agentic 工作流、自主软件 engineering 和科学 research 设计。它继承了最初的 Fable 5 模型,同时保持了相同的基本 token 定价,并将 prompt cache 读取速率降低了 75%。该 model 建立在 adaptive thinking 架构之上,使其能够在百万 token 窗口内跨多步骤执行循环进行规划、执行和自我纠正,而不会丢失 context。

核心能力

在架构上,Fable 5.1 针对 agentic 持久性进行了优化,使其能够处理跨越数小时或数天执行的任务,例如全栈应用程序开发或多阶段科学数据建模。它结合了精炼的安全防护,将技术安全和生物学任务上的误报拒绝率降低了高达 60-85%。

目标工作负载

该模型专为构建自主 agent 的开发人员和企业而设计,在多文件重构和专家级推理方面的可靠性是其主要需求。原生视觉支持还允许它直接解析代码库内的技术图表、UI 模型和嵌套表格图表。

Claude Fable 5.1

Claude Fable 5.1 的使用案例

发现使用 Claude Fable 5.1 获得出色效果的不同方式。

自主软件工程

导航复杂的跨目录仓库以定位架构缺陷并连贯地执行多文件重构。

科学数据建模

自动化解析原始观测数据、微分方程拟合以及用于研究的神经网络训练。

防御性安全审计

检查内部 API 和源代码以识别软件漏洞并建议修复补丁。

视觉 UI/UX 原型设计

从文本描述中生成功能性的 3D 模拟、交互式 WebGL 场景和生产就绪的前端代码。

多源知识综合

摄入海量的 SEC 申报文件、文字记录和图表数据集,以构建详细的财务或分析估值模型。

长周期 Agent 编排

充当异步工作流的主控制器,将子任务委托给工具并在数十个轮次中监控进度。

优势

局限性

精英科学推理: 在 GPQA Diamond 上达到 93.7%,在 Terminal-Bench 上达到 55.8%,用于研究和技术自动化。
非缓存定价较高: $10/M 输入和 $50/M 输出 tokens 的基础定价使得未缓存的单次调用成本高昂。
海量 Agent 能力: 提供 1,000,000 token 的 context window 和 128,000 的最大输出 tokens,用于长时间的自主执行。
Token 密集的内部起草: 扩展的 adaptive thinking 周期在最终输出前消耗大量的输出 tokens。
优化的缓存经济性: $0.25/M tokens 的 cache read 定价降低了多步骤 prompt 循环的总运营成本。
无原生视频支持: 模型处理静态图像和文档,但无法摄入原始视频文件。
精炼的安全系统: 在防御性网络安全和漏洞任务中,将良性误报拦截减少多达 60%。
全文件替换倾向: 对于简单的编辑,倾向于重写整个源文件,除非明确指示输出 diffs。

API快速入门

anthropic/claude-fable-5.1

查看文档
anthropic SDK
import Anthropic from '@anthropic-ai/sdk';

const anthropic = new Anthropic({
  apiKey: process.env.ANTHROPIC_API_KEY,
});

const msg = await anthropic.messages.create({
  model: "claude-fable-5.1",
  max_tokens: 1024,
  messages: [{ role: "user", content: "Analyze this codebase for architectural flaws." }],
});

console.log(msg.content[0].text);

安装SDK并在几分钟内开始进行API调用。

人们对 Claude Fable 5.1 的评价

看看社区对 Claude Fable 5.1 的看法

Fable 5.1 是第一个真正让人感觉像是用于长时间重构的 co-pilot 的模型,而且在半途不会迷失方向。
DevOpsGuru
reddit
虽然先前的模型在工作时间越长时越难以跟上,但 Fable 5.1 在漫长的多步骤任务中依然保持可读性。
Craig Falls
hackernews
Claude Fable 5.1 现在已在 Cursor 中可用!这是我们在 CursorBench 3.2 上运行过的能力最强的模型,得分为 73.4%。
cursor_ai
twitter
Anthropic 的 Mythos 级别模型终于攻克了长 agent 问题。Fable 5.1 的持久性无与伦比。
HackernewsUser42
hackernews
Fable 5.1 太疯狂了,兄弟们。它一次性耗尽了我的会话使用限制,但干净利落地完成了整个迁移。
TechLeadDev
twitter
prompt cache 的降价使得在生产环境中运行持续的测试-修复循环真正切实可行。
VibeCoder
youtube

关于 Claude Fable 5.1 的视频

观看关于 Claude Fable 5.1 的教程、评测和讨论

对于典型工作负载,Fable 5.1 的成本预计比 Fable 5 低约 25%

对于高度 agentic 的工作,节省通常会更大,高达大约 50%

Claude Mythos 5.1 与 Fable 5.1 相同,但它为经过审核的个人和组织提供了更宽松的安全防护

阴影看起来更加真实,物理效果感觉更真实

被要求这样做。有趣的是,我们一直在称赞它无需你要求就能把所有小细节都处理好的愿望或能力。这是它出于某种原因表现没那么好的地方之一。

如果你在打字,请静音。它恰好比 Opus 5 的 token 效率高出约 50%。它也比 Opus 5 更快。所以,如果

打磨得很完善。嗯,我认为就未来而言,我们将在如何使其变得更简单方面探讨更多。我认为

不仅仅是提示词

用以下方式提升您的工作流程 AI自动化

Automatio结合AI代理、网页自动化和智能集成的力量,帮助您在更短的时间内完成更多工作。

AI代理
网页自动化
智能工作流

Claude Fable 5.1专业提示

专家提示助您充分利用Claude Fable 5.1。

最大化缓存节省

将系统指令和大型仓库树保持在 prompt 的顶部,以利用 $0.25/M 的 cache read 定价。

控制执行强度

使用 effort 参数在用于标准任务的“low”和用于复杂调试的“high”之间切换,以平衡成本和速度。

提示进行针对性修改

明确指示模型提供统一的 diff 或针对性的行更改,以防止其重写整个大文件。

推动自主持久性

在 agentic 循环中,包含提示模型直接完成多步骤任务而不是暂停以获取权限的指令。

用户评价

用户怎么说

加入数千名已改变工作流程的满意用户

Jonathan Kogan

Jonathan Kogan

Co-Founder/CEO, rpatools.io

Automatio is one of the most used for RPA Tools both internally and externally. It saves us countless hours of work and we realized this could do the same for other startups and so we choose Automatio for most of our automation needs.

Mohammed Ibrahim

Mohammed Ibrahim

CEO, qannas.pro

I have used many tools over the past 5 years, Automatio is the Jack of All trades.. !! it could be your scraping bot in the morning and then it becomes your VA by the noon and in the evening it does your automations.. its amazing!

Ben Bressington

Ben Bressington

CTO, AiChatSolutions

Automatio is fantastic and simple to use to extract data from any website. This allowed me to replace a developer and do tasks myself as they only take a few minutes to setup and forget about it. Automatio is a game changer!

Sarah Chen

Sarah Chen

Head of Growth, ScaleUp Labs

We've tried dozens of automation tools, but Automatio stands out for its flexibility and ease of use. Our team productivity increased by 40% within the first month of adoption.

David Park

David Park

Founder, DataDriven.io

The AI-powered features in Automatio are incredible. It understands context and adapts to changes in websites automatically. No more broken scrapers!

Emily Rodriguez

Emily Rodriguez

Marketing Director, GrowthMetrics

Automatio transformed our lead generation process. What used to take our team days now happens automatically in minutes. The ROI is incredible.

Jonathan Kogan

Jonathan Kogan

Co-Founder/CEO, rpatools.io

Automatio is one of the most used for RPA Tools both internally and externally. It saves us countless hours of work and we realized this could do the same for other startups and so we choose Automatio for most of our automation needs.

Mohammed Ibrahim

Mohammed Ibrahim

CEO, qannas.pro

I have used many tools over the past 5 years, Automatio is the Jack of All trades.. !! it could be your scraping bot in the morning and then it becomes your VA by the noon and in the evening it does your automations.. its amazing!

Ben Bressington

Ben Bressington

CTO, AiChatSolutions

Automatio is fantastic and simple to use to extract data from any website. This allowed me to replace a developer and do tasks myself as they only take a few minutes to setup and forget about it. Automatio is a game changer!

Sarah Chen

Sarah Chen

Head of Growth, ScaleUp Labs

We've tried dozens of automation tools, but Automatio stands out for its flexibility and ease of use. Our team productivity increased by 40% within the first month of adoption.

David Park

David Park

Founder, DataDriven.io

The AI-powered features in Automatio are incredible. It understands context and adapts to changes in websites automatically. No more broken scrapers!

Emily Rodriguez

Emily Rodriguez

Marketing Director, GrowthMetrics

Automatio transformed our lead generation process. What used to take our team days now happens automatically in minutes. The ROI is incredible.

相关 AI Models

openai

GPT-5.5

OpenAI

GPT-5.5 is OpenAI's flagship frontier model with a 1M context window and five reasoning effort levels, optimized for autonomous agentic workflows and coding.

1M context
$5.00/$30.00/1M
xai

Grok-3

xAI

Grok-3 is xAI's flagship reasoning model, featuring deep logic deduction, a 128k context window, and real-time integration with X for live research and coding.

1M context
$3.00/$15.00/1M
moonshot

Kimi K3

Moonshot

Kimi K3 is Moonshot AI's 2.8T MoE model with a 1M token context window, native multimodal vision, and frontier-tier coding performance for complex agents.

1M context
$3.00/$15.00/1M
google

Gemini 3.1 Flash Live Preview

Google

Gemini 3.1 Flash Live Preview is Google's ultra-low-latency, audio-to-audio model featuring a 131K context window, high-fidelity multimodal reasoning, and...

131K context
$0.75/$4.50/1M
anthropic

Claude Opus 4.7

Anthropic

Claude Opus 4.7 is Anthropic's flagship model with a 1-million-token context, adaptive reasoning, and 3.3x vision resolution for enterprise-scale agents.

1M context
$5.00/$25.00/1M
openai

GPT-5.2 Pro

OpenAI

GPT-5.2 Pro is OpenAI's 2025 flagship reasoning model featuring Extended Thinking for SOTA performance in mathematics, coding, and expert knowledge work.

400K context
$21.00/$168.00/1M
google

Gemini 3.1 Pro

Google

Gemini 3.1 Pro is Google's elite multimodal model featuring the DeepThink reasoning engine, a 1M+ context window, and industry-leading ARC-AGI logic scores.

1M context
$2.00/$12.00/1M
alibaba

Qwen 3.7 Max

alibaba

Qwen 3.7 Max is Alibaba’s flagship AI model for deep reasoning and autonomous agent tasks, featuring a 256k context window and top-tier coding performance.

256K context
$1.20/$6.00/1M

关于Claude Fable 5.1的常见问题

查找关于Claude Fable 5.1的常见问题答案