All Comparisons

GPT-4o vs Claude Sonnet 4

GPT-4o
OpenAI 🇺🇸
58.1
#145
high
$2.50/M in · $10.00/M out
128K context
Claude Sonnet 4
Anthropic 🇺🇸
68.3
#93
high
$3.00/M in · $15.00/M out
1M context
Most Affordable
GPT-4o
$2.50/M input
Largest Context
Claude Sonnet 4
1M tokens
Best Benchmark Score
Claude Sonnet 4
68.3/100

Model Specifications

SpecGPT-4oClaude Sonnet 4
ProviderOpenAIAnthropic
Release DateMay 2024May 2025
Knowledge Cutoff2023-10-312025-01-31
Parametersunknownunknown
Context Window128K1M
Max Output16.4K64K
Open SourceNoNo
Licenseproprietaryproprietary
TokenizerGPTClaude
Modalitytext+image+file->texttext+image+file->text
Reasoning ModelNoNo
ModeratedYesYes

Pricing Comparison

Price (per 1M tokens)GPT-4oClaude Sonnet 4
Input$2.50$3.00
Output$10.00$15.00
Cache Read$1.25$0.30
Cache Write$3.75

API Provider Pricing

Capabilities

FeatureGPT-4oClaude Sonnet 4
Text Input
Image Input (Vision)
Audio Input
Video Input
File Input
Image Output
Audio Output
Tool Use
Structured Output (JSON)
Streaming
Reasoning Tokens
Web Search
Temperature Control
Top-P Sampling
Stop Sequences
Seed (Reproducibility)
Log Probabilities

Category Comparison

Benchmark-by-Benchmark

general

BenchmarkGPT-4oClaude Sonnet 4
MMLU-Pro
74.5
80.2
Chatbot Arena Elo
1285
1310
IFEval
83.6
90.1

coding

BenchmarkGPT-4oClaude Sonnet 4
HumanEval
90.2
93.8
SWE-bench Verified
33.2
53.2

math

BenchmarkGPT-4oClaude Sonnet 4
GSM8K
95.8
97.2
MATH
76.6
81.4
AIME 2024
13.4
24.5

reasoning

BenchmarkGPT-4oClaude Sonnet 4
BigBench-Hard
87.3
90.8
ARC-Challenge
96.4
GPQA Diamond
53.6
70.5

language

BenchmarkGPT-4oClaude Sonnet 4
WinoGrande
85.7
HellaSwag
96.4

multimodal

BenchmarkGPT-4oClaude Sonnet 4
MMMU
69.1
72.5

safety

BenchmarkGPT-4oClaude Sonnet 4
TruthfulQA
73.5

Category Winners

coding
Claude Sonnet 4
53.9
math
Claude Sonnet 4
66.8
reasoning
Claude Sonnet 4
70.5
general
Claude Sonnet 4
90.0
language
GPT-4o
79.9
multimodal
Claude Sonnet 4
62.3
safety
GPT-4o
94.3

Frequently Asked Questions