ModelDial · Model results

k3

One configuration tested. Explore its results and compare with other models.

Highest overall score74.5

k3 · high

Backend & testing
70
Frontend & interaction
85
Knowledge & reasoning
70
Evaluation time
158m 17s
Reference cost
$3.28

Totals for all three evaluations at this configuration; cost estimated from API usage.

Overall rank 12 / 33 models · Best tested effort

Backend tested · 9/10/2026

Published summary scores use the same effort. Axes may be evaluated on different dates; a backend date refers only to backend testing.

View task coverage and scoring →

Tested configuration

Scores, evaluation time and reference cost for this configuration.

Swipe to see all results →

Model / effortOverallBackend & testingFrontend & interactionKnowledge & reasoningEvaluation timeReference costCompare
high74.5708570158m 17s$3.28

How to read these results

The summary uses the same configuration as the highest overall score. A complete overall score requires all three axes; missing results appear as dashes.

Data snapshot published 10/1/2026