gpt-6.1-sol
gpt-6.1-sol has 5 tested configurations. Its highest published overall score is 93.8 at max. See the axis results, time and cost below.
Overall score93.8
gpt-6.1-sol · max
- Backend & testing
- 95
- Frontend & interaction
- 96
- Knowledge & reasoning
- 90
- Evaluation time
- 62m 9s
- Reference cost
- $1.69
All three evaluations · API reference cost
What changes with more effort
Compare scores, cost and time at neighboring tested efforts.
xhigh max
- Overall score
- +4.1 pts 89.7 → 93.8
- Reference cost
- +$0.071 $1.62 → $1.69
- Evaluation time
- −16.8% 74m 41s → 62m 9s
Results by reasoning effort
Swipe to see all results →
| Model / effort | Overall | Backend & testing | Frontend & interaction | Knowledge & reasoning | Evaluation time | Reference cost | Compare |
|---|---|---|---|---|---|---|---|
| low | 87.3 | 84 | 94 | 85 | 24m 22s | $1.33 | |
| medium | 87.5 | 89 | 83 | 90 | 28m 54s | $1.41 | |
| high | 90.7 | 91 | 96 | 85 | 42m 39s | $1.56 | |
| xhigh | 89.7 | 90 | 94 | 85 | 74m 41s | $1.62 | |
| max | 93.8 | 95 | 96 | 90 | 62m 9s | $1.69 |
Model and effort comparisons
GPT-6.1 Sol vs Claude Opus 5.5
Compare each model at its highest-scoring tested overall effort. Start with the capability relevant to your task, then compare completion time and reference cost.
GPT-6.1 Sol vs GPT-6 Sol
Compare the two Sol generations at their highest-scoring tested overall efforts, including whether score changes come with time or cost changes.
Data snapshot published · 10/2/2026Scoring & cost notes