All Comparisons

Claude Opus 4 vs GPT-4o

Claude Opus 4
Anthropic πŸ‡ΊπŸ‡Έ
75.1
#52β€”
high
$15.00/M in Β· $75.00/M out
200K context
GPT-4o
OpenAI πŸ‡ΊπŸ‡Έ
58.3
#136β€”
high
$2.50/M in Β· $10.00/M out
128K context
Most Affordable
GPT-4o
$2.50/M input
Largest Context
Claude Opus 4
200K tokens
Best Benchmark Score
Claude Opus 4
75.1/100

Model Specifications

SpecClaude Opus 4GPT-4o
ProviderAnthropicOpenAI
Release DateMay 2025May 2024
Knowledge CutoffMar 2025Oct 2023
Parametersunknownunknown
Context Window200K128K
Max Outputβ€”β€”
Open SourceNoNo
Licenseproprietaryproprietary
Tokenizerβ€”β€”
Modalityβ€”β€”
Reasoning ModelNoNo
ModeratedNoNo

Pricing Comparison

Price (per 1M tokens)Claude Opus 4GPT-4o
Input$15.00$2.50
Output$75.00$10.00

API Provider Pricing

Capabilities

FeatureClaude Opus 4GPT-4o
Text Input
Image Input (Vision)
Audio Input
Video Input
File Input
Image Output
Audio Output
Tool Use
Structured Output (JSON)
Streaming
Reasoning Tokens
Web Search
Temperature Control
Top-P Sampling
Stop Sequences
Seed (Reproducibility)
Log Probabilities

Category Comparison

Benchmark-by-Benchmark

general

BenchmarkClaude Opus 4GPT-4o
Chatbot Arena Elo
1365
1285
IFEval
91.5
83.6
MMLU-Pro
83.5
74.5

coding

BenchmarkClaude Opus 4GPT-4o
HumanEval
95.1
90.2
SWE-bench Verified
55.8
33.2

math

BenchmarkClaude Opus 4GPT-4o
MATH
85.2
76.6
GSM8K
97.8
95.8
AIME 2024
35
13.4

reasoning

BenchmarkClaude Opus 4GPT-4o
GPQA Diamond
74.8
53.6
BigBench-Hard
93.2
87.3
ARC-Challengeβ€”
96.4

multimodal

BenchmarkClaude Opus 4GPT-4o
MMMU
74.2
69.1

language

BenchmarkClaude Opus 4GPT-4o
HellaSwagβ€”
96.4
WinoGrandeβ€”
85.7

safety

BenchmarkClaude Opus 4GPT-4o
TruthfulQAβ€”
73.5

Category Winners

coding
Claude Opus 4
58.1
math
Claude Opus 4
75.5
reasoning
Claude Opus 4
79.7
general
Claude Opus 4
96.0
language
GPT-4o
79.9
multimodal
Claude Opus 4
67.8
safety
GPT-4o
94.3

Compare More Models

Frequently Asked Questions