All Comparisons

DeepSeek-R1 vs GPT-4.1

DeepSeek-R1
DeepSeek πŸ‡¨πŸ‡³
GPT-4.1
OpenAI πŸ‡ΊπŸ‡Έ
32.8
#238β€”
high
$2.00/M in Β· $8.00/M out
1.0M context
Most Affordable
GPT-4.1
$2.00/M input
Largest Context
GPT-4.1
1.0M tokens
Best Benchmark Score
GPT-4.1
32.8/100

Model Specifications

SpecDeepSeek-R1GPT-4.1
ProviderDeepSeekOpenAI
Release DateJan 2025Apr 2025
Knowledge Cutoffβ€”2024-06-01
Parameters671000000000BUndisclosed
Context Windowβ€”1.0M
Max Outputβ€”β€”
Open SourceYesNo
Licensemitproprietary
Tokenizerβ€”β€”
Modalityβ€”β€”
Reasoning ModelYesNo
ModeratedNoNo

Pricing Comparison

Price (per 1M tokens)DeepSeek-R1GPT-4.1
Inputβ€”$2.00
Outputβ€”$8.00

API Provider Pricing

Capabilities

FeatureDeepSeek-R1GPT-4.1
Text Input
Image Input (Vision)
Audio Input
Video Input
File Input
Image Output
Audio Output
Tool Use
Structured Output (JSON)
Streaming
Reasoning Tokens
Web Search
Temperature Control
Top-P Sampling
Stop Sequences
Seed (Reproducibility)
Log Probabilities

Category Comparison

Benchmark-by-Benchmark

general

BenchmarkDeepSeek-R1GPT-4.1
MMMLUβ€”
87.3

coding

BenchmarkDeepSeek-R1GPT-4.1
SWE-bench Verifiedβ€”
54.6

math

BenchmarkDeepSeek-R1GPT-4.1
AIME 2025β€”
46.4

reasoning

BenchmarkDeepSeek-R1GPT-4.1
GPQA Diamondβ€”
66.3
Humanity's Last Examβ€”
5.4

multimodal

BenchmarkDeepSeek-R1GPT-4.1
MMMUβ€”
74.8
CharXiv Reasoningβ€”
56.7

agent

BenchmarkDeepSeek-R1GPT-4.1
TAU-Bench Retailβ€”
68

Category Winners

coding
GPT-4.1
25.1
math
GPT-4.1
6.0
reasoning
GPT-4.1
27.5
general
GPT-4.1
66.0
multimodal
GPT-4.1
40.2
agent
GPT-4.1
50.7

Compare More Models

Frequently Asked Questions