All Comparisons

Claude Opus 4 vs Grok-3

Claude Opus 4
Anthropic 🇺🇸
74.8
#59
high
$15.00/M in · $75.00/M out
200K context
X
Grok-3
xAI 🇺🇸
77.5
#44
low
$3.00/M in · $15.00/M out
128K context
Most Affordable
Grok-3
$3.00/M input
Largest Context
Claude Opus 4
200K tokens
Best Benchmark Score
Grok-3
77.5/100

Model Specifications

SpecClaude Opus 4Grok-3
ProviderAnthropicxAI
Release DateMay 2025Feb 2025
Knowledge CutoffMar 20252024-11-17
ParametersunknownUndisclosed
Context Window200K128K
Max Output
Open SourceNoNo
Licenseproprietaryproprietary
Tokenizer
Modality
Reasoning ModelNoNo
ModeratedNoNo

Pricing Comparison

Price (per 1M tokens)Claude Opus 4Grok-3
Input$15.00$3.00
Output$75.00$15.00

API Provider Pricing

Capabilities

FeatureClaude Opus 4Grok-3
Text Input
Image Input (Vision)
Audio Input
Video Input
File Input
Image Output
Audio Output
Tool Use
Structured Output (JSON)
Streaming
Reasoning Tokens
Web Search
Temperature Control
Top-P Sampling
Stop Sequences
Seed (Reproducibility)
Log Probabilities

Category Comparison

Benchmark-by-Benchmark

general

BenchmarkClaude Opus 4Grok-3
Chatbot Arena Elo
1365
IFEval
91.5
MMLU-Pro
83.5

coding

BenchmarkClaude Opus 4Grok-3
SWE-bench Verified
55.8
HumanEval
95.1

math

BenchmarkClaude Opus 4Grok-3
GSM8K
97.8
AIME 2024
35
MATH
85.2
AIME 2025
93.3

reasoning

BenchmarkClaude Opus 4Grok-3
GPQA Diamond
74.8
84.6
BigBench-Hard
93.2

multimodal

BenchmarkClaude Opus 4Grok-3
MMMU
74.2
78

Category Winners

coding
Claude Opus 4
58.0
math
Claude Opus 4
75.5
reasoning
X
Grok-3
80.0
general
Claude Opus 4
96.0
multimodal
X
Grok-3
75.9

Compare More Models

Frequently Asked Questions