All Comparisons

Claude Sonnet 4 vs DeepSeek-R1
Claude Sonnet 4
Anthropic πΊπΈ
68.3
#93β
high$3.00/M in Β· $15.00/M out
200K context
DeepSeek-R1
DeepSeek π¨π³
Most Affordable
Claude Sonnet 4
$3.00/M input
Largest Context
Claude Sonnet 4
200K tokens
Best Benchmark Score
Claude Sonnet 4
68.3/100
Model Specifications
| Spec | Claude Sonnet 4 | DeepSeek-R1 |
|---|---|---|
| Provider | Anthropic | DeepSeek |
| Release Date | May 2025 | Jan 2025 |
| Knowledge Cutoff | Mar 2025 | β |
| Parameters | unknown | 671000000000B |
| Context Window | 200K | β |
| Max Output | β | β |
| Open Source | No | Yes |
| License | proprietary | mit |
| Tokenizer | β | β |
| Modality | β | β |
| Reasoning Model | No | Yes |
| Moderated | No | No |
Pricing Comparison
| Price (per 1M tokens) | Claude Sonnet 4 | DeepSeek-R1 |
|---|---|---|
| Input | $3.00 | β |
| Output | $15.00 | β |
API Provider Pricing
Capabilities
| Feature | Claude Sonnet 4 | DeepSeek-R1 |
|---|---|---|
| Text Input | ||
| Image Input (Vision) | ||
| Audio Input | ||
| Video Input | ||
| File Input | ||
| Image Output | ||
| Audio Output | ||
| Tool Use | ||
| Structured Output (JSON) | ||
| Streaming | ||
| Reasoning Tokens | ||
| Web Search | ||
| Temperature Control | ||
| Top-P Sampling | ||
| Stop Sequences | ||
| Seed (Reproducibility) | ||
| Log Probabilities |
Category Comparison
Benchmark-by-Benchmark
general
| Benchmark | Claude Sonnet 4 | DeepSeek-R1 |
|---|---|---|
| IFEval | 90.1 | β |
| Chatbot Arena Elo | 1310 | β |
| MMLU-Pro | 80.2 | β |
coding
| Benchmark | Claude Sonnet 4 | DeepSeek-R1 |
|---|---|---|
| SWE-bench Verified | 53.2 | β |
| HumanEval | 93.8 | β |
reasoning
| Benchmark | Claude Sonnet 4 | DeepSeek-R1 |
|---|---|---|
| BigBench-Hard | 90.8 | β |
| GPQA Diamond | 70.5 | β |
multimodal
| Benchmark | Claude Sonnet 4 | DeepSeek-R1 |
|---|---|---|
| MMMU | 72.5 | β |
Category Winners
coding
53.9
math
66.8
reasoning
70.5
general
90.0
multimodal
62.3