THE SCOREBOARD FOR WHAT'S NEXTAI + QUANTUM, SCORED. COMPARED. EXPLAINED.
AIQUANTUMSCORE.

VALUE FRONTIER

Performance
per dollar.

Only models with a verified category score and current USD API price appear. Each category stands alone; there is no universal winner.

overall evidence vs standard API price

Gemini 3.8 FlashGrok 4.6Grok 4.5GPT-6 SolClaude Sonnet 5.5GPT-5.6 SolClaude Opus 5.5GPT-6 AstraBlended input/output USD per 1M tokens →

Green points are efficient within this eligible cohort: no other plotted model has both a lower price and an equal or higher score. Prices use the stated 75/25 input/output mix.

ModelCategory scoreBlended priceConfidenceFrontier
Gemini 3.8 Flash51$1.50Moderate confidence · 3 independent domainsEfficient in this cohort
Grok 4.652$3.00Moderate confidence · 2 independent domainsEfficient in this cohort
Grok 4.544$3.00Moderate confidence · 2 independent domains—
GPT-6 Sol48$4.00Moderate confidence · 3 independent domains—
Claude Sonnet 5.588$4.00Moderate confidence · 2 independent domainsEfficient in this cohort
GPT-5.6 Sol53$8.00Moderate confidence · 3 independent domains—
Claude Opus 5.586$8.00Moderate confidence · 3 independent domains—
GPT-6 Astra60$20.00Moderate confidence · 4 independent domains—

Median standard API price

2026-09-292026-09-29
  • 2026-09-29: 3 USD / 1M · Median of 24 priced models

One observation; no trend can be inferred.

Qualified performance per dollar

No dated qualified overall scores yet; no trend is shown.

ARC-AGI-3 Standard benchmark frontier

2026-09-292026-09-29
  • 2026-09-29: 62.713 % · gpt-6-astra · ARC-AGI-3 Semi-Private Standard harness best effort 3 Semi-Private · source ↗

One observation; no trend can be inferred.