All Comparisons

Claude Sonnet 4 vs GPT-4.1

Claude Sonnet 4
Anthropic πŸ‡ΊπŸ‡Έ
68.5
#83β€”
high
$3.00/M in Β· $15.00/M out
200K context
GPT-4.1
OpenAI πŸ‡ΊπŸ‡Έ
32.8
#238β€”
high
$2.00/M in Β· $8.00/M out
1.0M context
Most Affordable
GPT-4.1
$2.00/M input
Largest Context
GPT-4.1
1.0M tokens
Best Benchmark Score
Claude Sonnet 4
68.5/100

Model Specifications

SpecClaude Sonnet 4GPT-4.1
ProviderAnthropicOpenAI
Release DateMay 2025Apr 2025
Knowledge CutoffMar 20252024-06-01
ParametersunknownUndisclosed
Context Window200K1.0M
Max Outputβ€”β€”
Open SourceNoNo
Licenseproprietaryproprietary
Tokenizerβ€”β€”
Modalityβ€”β€”
Reasoning ModelNoNo
ModeratedNoNo

Pricing Comparison

Price (per 1M tokens)Claude Sonnet 4GPT-4.1
Input$3.00$2.00
Output$15.00$8.00

API Provider Pricing

Capabilities

FeatureClaude Sonnet 4GPT-4.1
Text Input
Image Input (Vision)
Audio Input
Video Input
File Input
Image Output
Audio Output
Tool Use
Structured Output (JSON)
Streaming
Reasoning Tokens
Web Search
Temperature Control
Top-P Sampling
Stop Sequences
Seed (Reproducibility)
Log Probabilities

Category Comparison

Benchmark-by-Benchmark

general

BenchmarkClaude Sonnet 4GPT-4.1
IFEval
90.1
β€”
MMLU-Pro
80.2
β€”
Chatbot Arena Elo
1310
β€”
MMMLUβ€”
87.3

coding

BenchmarkClaude Sonnet 4GPT-4.1
HumanEval
93.8
β€”
SWE-bench Verified
53.2
54.6

math

BenchmarkClaude Sonnet 4GPT-4.1
MATH
81.4
β€”
GSM8K
97.2
β€”
AIME 2024
24.5
β€”
AIME 2025β€”
46.4

reasoning

BenchmarkClaude Sonnet 4GPT-4.1
BigBench-Hard
90.8
β€”
GPQA Diamond
70.5
66.3
Humanity's Last Examβ€”
5.4

multimodal

BenchmarkClaude Sonnet 4GPT-4.1
MMMU
72.5
74.8
CharXiv Reasoningβ€”
56.7

agent

BenchmarkClaude Sonnet 4GPT-4.1
TAU-Bench Retailβ€”
68

Category Winners

coding
Claude Sonnet 4
54.0
math
Claude Sonnet 4
66.8
reasoning
Claude Sonnet 4
70.9
general
Claude Sonnet 4
90.0
multimodal
Claude Sonnet 4
63.3
agent
GPT-4.1
50.7

Compare More Models

Frequently Asked Questions