claude-fable-5-1
Compare tested reasoning efforts on private datasets. Scores and costs come from their recorded evaluations.
Highest overall score89.0
claude-fable-5-1 · max
- Backend & testing
- 83
- Frontend & interaction
- 91
- Knowledge & reasoning
- 95
What changes with more effort
Compare scores, cost and time at neighboring tested efforts.
xhigh max
- Overall score
- +1.1 pts 87.9 → 89.0
- Reference cost
- +$10.21 $15.94 → $26.15
- Evaluation time
- +42.3% 55m 35s → 79m 7s
Results by reasoning effort
Select two or three configurations to compare scores and costs.
Swipe to see all results →
| Model / effort | Overall | Backend & testing | Frontend & interaction | Knowledge & reasoning | Evaluation time | Reference cost | Compare |
|---|---|---|---|---|---|---|---|
| high | 80.8 | 70 | 86 | 90 | 66m 53s | $8.28 | |
| xhigh | 87.9 | 81 | 90 | 95 | 55m 35s | $15.94 | |
| max | 89.0 | 83 | 91 | 95 | 79m 7s | $26.15 |
How to read these results
The summary uses the same configuration as the highest overall score. A complete overall score requires all three axes; missing results appear as dashes.
Evaluation details · Published 9/30/2026