ModelDial · Model results

qwen/qwen3.8-27b

One configuration tested. Explore its results and compare with other models.

Published results2 / 3 axes

No overall score until all three axes are published.

qwen/qwen3.8-27b · default

Backend & testing
64
Frontend & interaction
87
Knowledge & reasoning
—
Evaluation time
49m 30s
Reference cost
—

Totals cover only the published evaluations at this configuration; cost estimated from API usage.

Backend tested · 9/28/2026

Published summary scores use the same effort. Axes may be evaluated on different dates; a backend date refers only to backend testing.

View task coverage and scoring →

Tested configuration

Scores, evaluation time and reference cost for this configuration.

Swipe to see all results →

Model / effortOverallBackend & testingFrontend & interactionKnowledge & reasoningEvaluation timeReference costCompare
default—6487—49m 30s—

How to read these results

The summary shows the published axes at this configuration. A missing result is not zero and does not imply an evaluation is in progress.

Data snapshot published 10/1/2026