Skip to content

Sonnet 4.5 vs GPT-6.1 Sol

Compare pricing, context, and capabilities side by side. Then give your models the same coding task and see how they perform.

Quick compare
Sonnet 4.5anthropic/claude-sonnet-4.5

This model is not listed in the current public catalog. Its selection is preserved; published specs are unavailable.

View model details
GPT-6.1 Solopenai/gpt-6.1-sol

This model is not listed in the current public catalog. Its selection is preserved; published specs are unavailable.

View model details
Sonnet 4.5GPT-6.1 Sol

Overview

AuthorAnthropic
AuthorOpenAI
Context lengthUnknown
Context lengthUnknown
ReasoningUnknown
ReasoningUnknown
Input modalitiesUnknown
Input modalitiesUnknown
Output modalitiesUnknown
Output modalitiesUnknown

Pricing

USD per million tokens
InputUnknown
InputUnknown
OutputUnknown
OutputUnknown

Features

Tool useUnknown
Tool useUnknown
Max output tokensUnknown
Max output tokensUnknown
Reasoning effortsUnknown
Reasoning effortsUnknown

Published benchmarks

Refreshed 2026-10-01 · Stale snapshot · Incomplete coverage
SWE-benchVerified (500 tasks)
71.4%
Evaluation details
Variant
claude-sonnet-4-5-20250929
Version
Verified (500 tasks)
Harness
mini-SWE-agent 2.0.0
Settings
One attempt; reasoning high
Evaluated
2026-02-17
Refreshed
2026-10-01
View published result ↗
SWE-benchVerified (500 tasks)
No reviewed result
Terminal-BenchCore 0.1.1
58.75%
Evaluation details
Variant
Claude 4.5 Sonnet (directory label; dated API variant not reported)
Version
Core 0.1.1
Harness
Factory Droid; commit d243b033c6ca0b15c95c0126993040955730d8d7
Settings
Published run tb_rc6_s45_1 only; one attempt across 80 tasks; not the five-run aggregate
Evaluated
2025-09-29
Refreshed
2026-10-01
View published result ↗
Terminal-BenchCore 0.1.1
No reviewed result
Artificial Analysis Intelligence IndexVersion not reported on leaderboard
No reviewed result
Artificial Analysis Intelligence IndexVersion not reported on leaderboard
52index
Evaluation details
Variant
GPT-6.1 Sol (max)
Version
Version not reported on leaderboard
Harness
Artificial Analysis
Settings
max reasoning; other settings not reported
Evaluated
Date unknown
Refreshed
2026-10-01
View published result ↗

Results are matched to exact variants. Different harnesses and reasoning settings are shown in evaluation details; these scores are separate from your Codz tasks.

Published rates are separate from your task costs. Unknown values are not supplied by the catalog.

Catalog refreshes every 6 hours

See what works on your code.

One prompt, the same commit, independent attempts.

Compare a task

Prefer local models? Compare in the macOS app