qwen/qwen3.8-27b
One configuration tested. Explore its results and compare with other models.
Published results2 / 3 axes
No overall score until all three axes are published.
qwen/qwen3.8-27b · default
- Backend & testing
- 64
- Frontend & interaction
- 87
- Knowledge & reasoning
- —
- Evaluation time
- 49m 30s
- Reference cost
- —
Totals cover only the published evaluations at this configuration; cost estimated from API usage.
Backend tested · 9/28/2026
Published summary scores use the same effort. Axes may be evaluated on different dates; a backend date refers only to backend testing.
View task coverage and scoring →Tested configuration
Scores, evaluation time and reference cost for this configuration.
Swipe to see all results →
| Model / effort | Overall | Backend & testing | Frontend & interaction | Knowledge & reasoning | Evaluation time | Reference cost | Compare |
|---|---|---|---|---|---|---|---|
| default | — | 64 | 87 | — | 49m 30s | — |
How to read these results
The summary shows the published axes at this configuration. A missing result is not zero and does not imply an evaluation is in progress.
Data snapshot published 10/1/2026