Claude Fable 5 (Thinking) vs Claude Opus 4.6 (Thinking) vs Spark 1.1
Side-by-side benchmark scores, pricing, and specifications
Specifications
| Specification | |||
|---|---|---|---|
| Provider | Anthropic | Anthropic | Meta |
| Variant | Thinking | 4.6 Thinking | Spark 1.1 |
| Input price | $10.00/1M | $5.00/1M | — |
| Output price | $50.00/1M | $25.00/1M | — |
| Context window | — | 1.0M | — |
| Benchmark | Comparison | Claude Fable 5 (Thinking) | Claude Opus 4.6 (Thinking) | Spark 1.1 |
|---|---|---|---|---|
CompositeQuality Score | 114.2%#3 | 99.9%#12 | 106.4%#6 | |
Human preferenceArena ELO | 1,507#1 | 1,505#2 | 1,495#5 | |
General reasoningLiveBench | 80.8%#3 | 74.5%#23 | 76.2%#15 | |
Scientific knowledgeGPQA Diamond | — | 91.3%#11 | — | |
Academic reasoningHLE | — | 40.0%#9 | — | |
Commonsense reasoningSimpleBench | 81.9%#1 | 67.6%#9 | — | |
MathematicsAIME 2025 | — | 95.6%#2 | — | |
Graduate scienceGSO | — | 37.3%#3 | — | |
Agentic codingSWE-Bench Verified | — | 80.8%#5 | — | |
Agentic tool useTau-Bench | — | 91.9%#1 | — | |
Agentic terminalTerminal-Bench | — | 65.4%#5 | — | |
Visual reasoningARC-AGI-2 | — | 68.8%#10 | — |
Scores represent the best available variant for each model. Higher is better unless otherwise noted. Bars show relative performance within each benchmark.