Claude Fable 5.1

Claude Fable 5.1 is a reasoning model from Anthropic in the Claude Fable family. 55 benchmarks count toward its score, in 7 categories.

availableShows if the model has enough results for an index.
IndexOverall score out of 100.81.4 ±2.7
CoverageShare of the index weight with results.95%
SpeedOutput tokens per second.49/s
Input / 1MUS dollars per 1M input tokens.$10 batch $5
Output / 1MUS dollars per 1M output tokens.$50 batch $25 US dollars per 1M output tokens in a batch.
ContextMaximum tokens in one request.1M
EloLMArena rating and rank.1508 (#1)

The index is a score out of 100. The ± range shows how much it can change. Batch work costs less.

5,783 votes. Elo shows what people prefer. It does not change the score.

CapabilitiesScore per category, out of 100.

Out of 100
AgenticMulti-step tasks with tools.
79.2
CodingCode writing and repair.
79.9
ReasoningLogic problems and puzzles.
84.8
MultimodalTasks with images and text.
74.8
KnowledgeFacts and expert knowledge.
79.5
MultilingualTasks in many languages.
N/A
InstructionTasks with strict rules in the prompt.
75.2
MathMath problems.
81.4

Results

55 counted
BenchmarkThe test name.CategoryThe capability that the test measures.ResultThe score from the publisher.IndexThis result as a score out of 100.RunThe settings of the run.DateDate of the result.Published byThe source of the result.
ProofBench v1.1Math100.0%89.921 Sept 2026Vals AI
OTIS Mock AIME 2024-2025Math100.0%67.4max effortEpoch AI
ARC-AGI-1 (semi-private)Reasoning97.5%73.2max effortARC Prize Foundation
LiveBench MathematicsMath97.0%77.5max effort25 Jun 2026LiveBench
Artificial Analysis GPQA DiamondKnowledge93.7%64.6Artificial Analysis
GPQA DiamondKnowledge93.4%64.51 Sept 2026Vals AI
Artificial Analysis Harvey LAB-AAAgentic93.0%76.8Artificial Analysis
MMLU ProKnowledge92.4%66.11 Sept 2026Vals AI
LiveBench ReasoningReasoning91.7%83.3max effort25 Jun 2026LiveBench
IOICoding90.8%85.221 Sept 2026Vals AI
MMMU ProMultimodal90.6%74.81 Sept 2026Vals AI
LiveCodeBenchCoding90.5%66.61 Sept 2026Vals AI
Vibe Code Bench v1.1Coding90.3%79.7OpenHands21 Sept 2026Vals AI
FrontierMath-Tiers-1-3-v2-PrivateMath90.2%82.2max effortEpoch AI
ARC-AGI-2 (semi-private)Reasoning90.0%85.1max effortARC Prize Foundation
LiveBench LanguageKnowledge89.5%78.5max effort25 Jun 2026LiveBench
FrontierMath-Tier-4-v2-PrivateMath87.8%90.1max effortEpoch AI
ProgramBench: Can Language Models Rebuild Programs From Scratch?Coding87.6%82.1John Yang et al.
LiveBench CodingCoding86.4%81.0max effort25 Jun 2026LiveBench
Artificial Analysis Long Context ReasoningReasoning85.3%67.3Artificial Analysis
Terminal-Bench 2.1Agentic85.0%74.221 Sept 2026Vals AI
Artificial Analysis Coding IndexCoding81.6%76.5Artificial Analysis
Toolathlon Verified Pass@3Agentic81.5%Anthropic
SWE-bench ProCoding81.2%82.6Xiang Deng et al.
LiveBench Data AnalysisReasoning80.3%67.5max effort25 Jun 2026LiveBench
Toolathlon-VerifiedAgentic77.8%75.4Moonshot AI
cursorBench32Coding73.4%81.1Benchmark authors
MirrorCodeCoding73.3%high effortEpoch AI
Toolathlon Verified Pass cubedAgentic73.1%Anthropic
LiveBench Instruction FollowingInstruction73.0%75.2max effort25 Jun 2026LiveBench
ApprenticeBench: end-to-end computer use, continual learning, and long-horizon agency on a real accounts-payable jobAgentic72.0%95.0NeoCognition
Medical Long Context Reasoning (MLCR-AA)Reasoning71.1%95.0Wisedocs and Artificial Analysis
SimpleQA VerifiedKnowledge70.8%87.1max effortEpoch AI
Furniture AssemblyReasoning70.0%89.4max effortEpoch AI
DeepSWEAgentic67.4%72.8Datacurve AI
Artificial Analysis Omniscience AccuracyKnowledge67.2%95.0Artificial Analysis
LiveBench Agentic CodingAgentic66.1%78.9max effort25 Jun 2026LiveBench
Humanity's Last ExamKnowledge65.0%83.8Center for AI Safety et al.
Artificial Analysis SciCodeCoding63.1%80.4Artificial Analysis
GDPval-AA normalizedAgentic61.7%83.0Artificial Analysis
SkillsBenchCoding61.6%75.2OpenHands11 Sept 2026Vals AI
Humanity's Last Exam without toolsKnowledge60.9%80.4OpenAI
Artificial Analysis AutomationBenchAgentic59.4%71.2Artificial Analysis
Artificial Analysis Humanity's Last ExamKnowledge59.1%88.6Artificial Analysis
Mystery Game PuzzlesReasoning58.0%95.0max effortEpoch AI
Artificial Analysis Agentic IndexAgentic58.0%84.4Artificial Analysis
Artificial Analysis AnalystAgentAgentic57.5%81.9Artificial Analysis
EBR-benchReasoning57.1%88.4max effortEpoch AI
FrontierSWE v2Coding56.3%85.6Proximal
Code MigrationCoding54.6%79.921 Sept 2026Vals AI
Artificial Analysis Intelligence IndexKnowledge53.4%89.1Artificial Analysis
Terminal-Bench-Science 0.1Agentic52.6%Terminal-Bench-Science Team
cursorBench40Coding51.8%Benchmark authors
Terminal-Bench 4.0Agentic49.5%94.321 Sept 2026Vals AI
Artificial Analysis Tau3-BankingAgentic47.2%76.7Artificial Analysis
Chess PuzzlesReasoning47.0%84.3max effortEpoch AI
Bug Hunt BenchCoding43.0%Pawel Huryn
OSWorld 2.0Agentic41.7%75.8Mengqi Yuan et al.
AutomationBenchAgentic31.4%67.7Moonshot AI
Critical Physics TasksReasoning29.7%95.0Artificial Analysis
Vibe Code Bench 1-100Coding28.0%82.6OpenHands16 Sept 2026Vals AI
Artificial Analysis GDP.pdfAgentic26.2%77.7Artificial Analysis
Toolathlon Verified average assistant turnsAgentic23.7%Anthropic
Agent Arena task outcomeAgentic19.889.7max effort15 Sept 2026LMArena
Agent Arena command recoveryAgentic12.681.6max effort15 Sept 2026LMArena
ProgramBenchCoding7.0%21 Sept 2026Vals AI
Agent Arena steerabilityAgentic3.971.7max effort15 Sept 2026LMArena
FrontierMath-ErdosMath0.0%max effortEpoch AI

55 benchmarks count, from 59 of 68 results. A grey row does not count. Too few models took that benchmark.

Sources

BenchLM benchmark aggregationUsed with attribution; per-benchmark results credited to their original authorsOpenRouter, collected directlyNo licence statedVals AI, collected directlyNo licence stated. Read from the public leaderboard and credited to Vals AIEpoch AI, collected directlyCC BY — free to use and redistribute with attributionARC Prize Foundation, collected directlyNo licence stated. Their terms ask for written permission before commercial useLiveBench, collected directlyNo licence stated for the leaderboard. The site repo serving these CSVs has no LICENSE; the harness repo carries upstream Apache-2.0 and MIT copies that cover code, not resultsLMArena, collected directlyCC BY 4.0 (lmarena-ai/leaderboard-dataset on Hugging Face)

Same level, lower price

Claude Opus 5.582.8 · $20

More from Anthropic

Claude Opus 5.582.8Claude Opus 577.9Claude Fable 577.5Claude Opus 4.871.1Claude Sonnet 567.6