THE SCOREBOARD FOR WHAT'S NEXTAI + QUANTUM, SCORED. COMPARED. EXPLAINED.
AIQUANTUMSCORE.

Anthropic / Claude

Claude Fable 5.1

Version claude-fable-5-1 · Released 2026-09-01 · Available

Compare this model ↗
Insufficient verified data73% coverage

Explore Anthropic history →

Coding43
Reasoning92
Math99
Research—
Agents40
Multimodal—
Speed—
Value23

01 / EVIDENCE

Benchmark breakdown

BenchmarkVersionMethodConfigurationResultTest dateVerifiedSource typeSource
Terminal-Bench 4.0 Vals mini-SWE-agent avg@3 fallbacks failed4.0Vals mini-SWE-agent avg@3; Anthropic fallbacks treated as failures—50 %2026-09-292026-09-29independent labView source ↗
ProofBench v1.1 Vals 100 proof tasks1.1Vals 100 proof tasks; ceiling effect at top scores—100 %2026-09-292026-09-29independent labView source ↗
ProgramBench Vals 200 public tasks raw pass rateVals 2026-09-27Vals 200 public tasks; raw pass rate; mini-SWE-agent—82.7 %2026-09-272026-09-29independent labView source ↗
LiveBench code_completion2026-06-25Public table; published model effort variantclaude-fable-5-1-max-effort82.609 %not published2026-09-30benchmark maintainerView source ↗
LiveBench code_generation2026-06-25Public table; published model effort variantclaude-fable-5-1-max-effort90.141 %not published2026-09-30benchmark maintainerView source ↗
LiveBench javascript2026-06-25Public table; published model effort variantclaude-fable-5-1-max-effort68.182 %not published2026-09-30benchmark maintainerView source ↗
LiveBench AMPS_Hard2026-06-25Public table; published model effort variantclaude-fable-5-1-max-effort99 %not published2026-09-30benchmark maintainerView source ↗
LiveBench olympiad2026-06-25Public table; published model effort variantclaude-fable-5-1-max-effort92.968 %not published2026-09-30benchmark maintainerView source ↗
LiveBench theory_of_mind2026-06-25Public table; published model effort variantclaude-fable-5-1-max-effort80.769 %not published2026-09-30benchmark maintainerView source ↗
LiveBench zebra_puzzle2026-06-25Public table; published model effort variantclaude-fable-5-1-max-effort100 %not published2026-09-30benchmark maintainerView source ↗

02 / COST

Price & speed

Input / 1M tokens $10.00

Output / 1M tokens $50.00

Cached input / 1M —

Batch input / 1M —

Batch output / 1M —

Speed No comparable observation yet

Price verified 2026-09-29

Price scope Base input/output rates; caching and batch tiers differ

Price source Official pricing ↗

02A / PERFORMANCE

Measured speed

No sourced speed measurement at a stated effort level yet.

03 / PROFILE

Capabilities

tool usevision

04 / CONTEXT

Strengths & limitations

Coverage: 73% · Limited evidence · 10 observations · 2 source domains (2 independent).

Why no Overall AIQuantumScore?

Qualified category evidence: coding, reasoning, math, agents. An overall score is withheld because the independent observation threshold has not been met. Evidence still missing or lacking comparable overlap: research, multimodal, speed.

See the benchmark coverage matrix →

05 / HISTORY

Version & price history

Version claude-fable-5-1 · released 2026-09-01.

2026-09-29: $10.00 input / $50.00 output per 1M tokens · Price source ↗

06 / SOURCES

Source records

DATED EVIDENCE

History

Standard API price history

2026-09-292026-09-29
  • 2026-09-29: 20 USD / 1M · 75% input / 25% output USD per 1M · source ↗

One observation; no trend can be inferred.

Terminal-Bench 4.0 Vals mini-SWE-agent avg@3 fallbacks failed / 4.0 / %

2026-09-292026-09-29
  • 2026-09-29: 50 result · Vals mini-SWE-agent avg@3; Anthropic fallbacks treated as failures · source ↗

One observation; no trend can be inferred.

ProofBench v1.1 Vals 100 proof tasks / 1.1 / %

2026-09-292026-09-29
  • 2026-09-29: 100 result · Vals 100 proof tasks; ceiling effect at top scores · source ↗

One observation; no trend can be inferred.

ProgramBench Vals 200 public tasks raw pass rate / Vals 2026-09-27 / %

2026-09-272026-09-27
  • 2026-09-27: 82.7 result · Vals 200 public tasks; raw pass rate; mini-SWE-agent · source ↗

One observation; no trend can be inferred.

LiveBench code_completion / 2026-06-25 / %

2026-09-302026-09-30
  • 2026-09-30: 82.609 result · Public table; published model effort variant · evaluation date not published; verification date shown · source ↗

One observation; no trend can be inferred.

LiveBench code_generation / 2026-06-25 / %

2026-09-302026-09-30
  • 2026-09-30: 90.141 result · Public table; published model effort variant · evaluation date not published; verification date shown · source ↗

One observation; no trend can be inferred.

LiveBench javascript / 2026-06-25 / %

2026-09-302026-09-30
  • 2026-09-30: 68.182 result · Public table; published model effort variant · evaluation date not published; verification date shown · source ↗

One observation; no trend can be inferred.

LiveBench AMPS_Hard / 2026-06-25 / %

2026-09-302026-09-30
  • 2026-09-30: 99 result · Public table; published model effort variant · evaluation date not published; verification date shown · source ↗

One observation; no trend can be inferred.

LiveBench olympiad / 2026-06-25 / %

2026-09-302026-09-30
  • 2026-09-30: 92.968 result · Public table; published model effort variant · evaluation date not published; verification date shown · source ↗

One observation; no trend can be inferred.

LiveBench theory_of_mind / 2026-06-25 / %

2026-09-302026-09-30
  • 2026-09-30: 80.769 result · Public table; published model effort variant · evaluation date not published; verification date shown · source ↗

One observation; no trend can be inferred.

LiveBench zebra_puzzle / 2026-06-25 / %

2026-09-302026-09-30
  • 2026-09-30: 100 result · Public table; published model effort variant · evaluation date not published; verification date shown · source ↗

One observation; no trend can be inferred.

coding category score history

2026-09-302026-09-30
  • 2026-09-30: 8 score · Scoring 3.0.0

One observation; no trend can be inferred.

math category score history

2026-09-302026-09-30
  • 2026-09-30: 100 score · Scoring 3.0.0

One observation; no trend can be inferred.

07 / EXPLORE

Similar models