ModelDial · Model results

qwen/qwen3.8-27b

One configuration tested. Explore its results and compare with other models.

Highest overall score—

qwen/qwen3.8-27b · default

Backend & testing
59
Frontend & interaction
87
Knowledge & reasoning
—
Evaluation time
—
Reference cost
—

Totals for all three evaluations at this configuration; cost estimated from API usage.

Backend tested · 9/28/2026

All three scores use the same effort. The date refers to backend testing only; evaluation dates may differ by axis.

View task coverage and scoring →

Tested configuration

Scores, evaluation time and reference cost for this configuration.

Swipe to see all results →

Model / effortOverallBackend & testingFrontend & interactionKnowledge & reasoningEvaluation timeReference costCompare
default—5987———

How to read these results

The summary uses the same configuration as the highest overall score. A complete overall score requires all three axes; missing results appear as dashes.

Data snapshot published 9/30/2026