- Published
- Aug 26, 2026, 11:00 PM UTC
- Question pack
- coding-fast-v4.10
- Coverage
- 5 measured configurations
- Batch
- snapshot-2026-08-26T23-00-00Z-r3
Measured configurations
| Configuration | Rank | Score | Elapsed | Reference cost | Route |
|---|---|---|---|---|---|
| GPT-5.6 Sol / Max | #1 | 85/100 | 44m 58s | $1.82 | Official login |
| GPT-5.6 Sol / XHigh | #3 | 83/100 | 45m 48s | $1.26 | Official login |
| GPT-5.6 Sol / High | #4 | 80/100 | 45m 2s | $0.96 | Official login |
| GPT-5.6 Sol / Medium | #9 | 68/100 | 43m 38s | $0.66 | Official login |
| GPT-5.6 Sol / Low | #13 | 65/100 | 39m 5s | $0.44 | Official login |
Five-question profile
| Question | GPT-5.6 Sol / Max | GPT-5.6 Sol / XHigh | GPT-5.6 Sol / High | GPT-5.6 Sol / Medium | GPT-5.6 Sol / Low |
|---|---|---|---|---|---|
| Black-box Regression Audit | 14/20 | 16/20 | 15/20 | 12/20 | 13/20 |
| Retry Planner Counterexamples | 20/20 | 18/20 | 17/20 | 17/20 | 14/20 |
| CI Adversarial Audit | 18/20 | 17/20 | 18/20 | 14/20 | 13/20 |
| Transaction Regression Design | 17/20 | 16/20 | 18/20 | 13/20 | 11/20 |
| Cache Regression Test Design | 16/20 | 16/20 | 12/20 | 12/20 | 14/20 |
What the effort spread can tell you
This family is measured at several effort levels through the recorded official-login route. The page keeps every effort separate so a strong Max result does not get attributed to Low, or vice versa.
Use the highest score and fastest completion as two different signals. If they point to different configurations, the choice depends on your quality floor rather than the family name.