All Comparisons

DeepSeek-V3 vs GPT-4.1

DeepSeek-V3
DeepSeek πŸ‡¨πŸ‡³
23.4
#265β€”
low
GPT-4.1
OpenAI πŸ‡ΊπŸ‡Έ
32.8
#238β€”
high
$2.00/M in Β· $8.00/M out
1.0M context
Most Affordable
GPT-4.1
$2.00/M input
Largest Context
GPT-4.1
1.0M tokens
Best Benchmark Score
GPT-4.1
32.8/100

Model Specifications

SpecDeepSeek-V3GPT-4.1
ProviderDeepSeekOpenAI
Release DateDec 2024Apr 2025
Knowledge Cutoffβ€”2024-06-01
Parameters671000000000BUndisclosed
Context Windowβ€”1.0M
Max Outputβ€”β€”
Open SourceYesNo
Licensemit_+_model_license_(commercial_use_allowed)proprietary
Tokenizerβ€”β€”
Modalityβ€”β€”
Reasoning ModelNoNo
ModeratedNoNo

Pricing Comparison

Price (per 1M tokens)DeepSeek-V3GPT-4.1
Inputβ€”$2.00
Outputβ€”$8.00

API Provider Pricing

Capabilities

FeatureDeepSeek-V3GPT-4.1
Text Input
Image Input (Vision)
Audio Input
Video Input
File Input
Image Output
Audio Output
Tool Use
Structured Output (JSON)
Streaming
Reasoning Tokens
Web Search
Temperature Control
Top-P Sampling
Stop Sequences
Seed (Reproducibility)
Log Probabilities

Category Comparison

Benchmark-by-Benchmark

general

BenchmarkDeepSeek-V3GPT-4.1
SimpleQA
24.9
β€”
MMMLUβ€”
87.3

coding

BenchmarkDeepSeek-V3GPT-4.1
SWE-bench Verified
42
54.6

reasoning

BenchmarkDeepSeek-V3GPT-4.1
GPQA Diamond
59.1
66.3
Humanity's Last Examβ€”
5.4

math

BenchmarkDeepSeek-V3GPT-4.1
AIME 2025β€”
46.4

multimodal

BenchmarkDeepSeek-V3GPT-4.1
MMMUβ€”
74.8
CharXiv Reasoningβ€”
56.7

agent

BenchmarkDeepSeek-V3GPT-4.1
TAU-Bench Retailβ€”
68

Category Winners

coding
GPT-4.1
25.1
math
GPT-4.1
6.0
reasoning
DeepSeek-V3
33.9
general
GPT-4.1
66.0
multimodal
GPT-4.1
40.2
agent
GPT-4.1
50.7

Compare More Models

Frequently Asked Questions