ModelDial · Model results

claude-fable-5-1

Compare tested reasoning efforts on private datasets. Scores and costs come from their recorded evaluations.

Highest overall score89.0

claude-fable-5-1 · max

Backend & testing
83
Frontend & interaction
91
Knowledge & reasoning
95

What changes with more effort

Compare scores, cost and time at neighboring tested efforts.

xhigh max
Overall score
+1.1 pts
87.9 89.0
Reference cost
+$10.21
$15.94 $26.15
Evaluation time
+42.3%
55m 35s 79m 7s
View full comparison →

Results by reasoning effort

Select two or three configurations to compare scores and costs.

Swipe to see all results →

Model / effortOverallBackend & testingFrontend & interactionKnowledge & reasoningEvaluation timeReference costCompare
high80.870869066m 53s$8.28
xhigh87.981909555m 35s$15.94
max89.083919579m 7s$26.15

How to read these results

The summary uses the same configuration as the highest overall score. A complete overall score requires all three axes; missing results appear as dashes.

Evaluation details · Published 9/30/2026