Sonnet 4.5 vs GPT-6.1 Sol
Compare pricing, context, and capabilities side by side. Then give your models the same coding task and see how they perform.
Sonnet 4.5
anthropic/claude-sonnet-4.5This model is not listed in the current public catalog. Its selection is preserved; published specs are unavailable.
View model detailsGPT-6.1 Sol
openai/gpt-6.1-solThis model is not listed in the current public catalog. Its selection is preserved; published specs are unavailable.
View model details| Sonnet 4.5 | GPT-6.1 Sol |
|---|---|
Overview | |
Author | Author |
Context lengthUnknown | Context lengthUnknown |
ReasoningUnknown | ReasoningUnknown |
Input modalitiesUnknown | Input modalitiesUnknown |
Output modalitiesUnknown | Output modalitiesUnknown |
PricingUSD per million tokens | |
InputUnknown | InputUnknown |
OutputUnknown | OutputUnknown |
Features | |
Tool useUnknown | Tool useUnknown |
Max output tokensUnknown | Max output tokensUnknown |
Reasoning effortsUnknown | Reasoning effortsUnknown |
Published benchmarksRefreshed 2026-10-01 · Stale snapshot · Incomplete coverage | |
SWE-benchVerified (500 tasks) 71.4% Evaluation details
| SWE-benchVerified (500 tasks) No reviewed result |
Terminal-BenchCore 0.1.1 58.75% Evaluation details
| Terminal-BenchCore 0.1.1 No reviewed result |
Artificial Analysis Intelligence IndexVersion not reported on leaderboard No reviewed result | Artificial Analysis Intelligence IndexVersion not reported on leaderboard 52index Evaluation details
|
Results are matched to exact variants. Different harnesses and reasoning settings are shown in evaluation details; these scores are separate from your Codz tasks.
Published rates are separate from your task costs. Unknown values are not supplied by the catalog.
Catalog refreshes every 6 hoursSee what works on your code.
One prompt, the same commit, independent attempts.
Prefer local models? Compare in the macOS app