All Comparisons

DeepSeek-V3 vs GPT-4o

DeepSeek-V3
DeepSeek 🇨🇳
23.4
#265
low
GPT-4o
OpenAI 🇺🇸
58.3
#136
high
$2.50/M in · $10.00/M out
128K context
Most Affordable
GPT-4o
$2.50/M input
Largest Context
GPT-4o
128K tokens
Best Benchmark Score
GPT-4o
58.3/100

Model Specifications

SpecDeepSeek-V3GPT-4o
ProviderDeepSeekOpenAI
Release DateDec 2024May 2024
Knowledge CutoffOct 2023
Parameters671000000000Bunknown
Context Window128K
Max Output
Open SourceYesNo
Licensemit_+_model_license_(commercial_use_allowed)proprietary
Tokenizer
Modality
Reasoning ModelNoNo
ModeratedNoNo

Pricing Comparison

Price (per 1M tokens)DeepSeek-V3GPT-4o
Input$2.50
Output$10.00

API Provider Pricing

Capabilities

FeatureDeepSeek-V3GPT-4o
Text Input
Image Input (Vision)
Audio Input
Video Input
File Input
Image Output
Audio Output
Tool Use
Structured Output (JSON)
Streaming
Reasoning Tokens
Web Search
Temperature Control
Top-P Sampling
Stop Sequences
Seed (Reproducibility)
Log Probabilities

Category Comparison

Benchmark-by-Benchmark

general

BenchmarkDeepSeek-V3GPT-4o
SimpleQA
24.9
MMLU-Pro
74.5
IFEval
83.6
Chatbot Arena Elo
1285

coding

BenchmarkDeepSeek-V3GPT-4o
SWE-bench Verified
42
33.2
HumanEval
90.2

reasoning

BenchmarkDeepSeek-V3GPT-4o
GPQA Diamond
59.1
53.6
BigBench-Hard
87.3
ARC-Challenge
96.4

math

BenchmarkDeepSeek-V3GPT-4o
AIME 2024
13.4
MATH
76.6
GSM8K
95.8

language

BenchmarkDeepSeek-V3GPT-4o
HellaSwag
96.4
WinoGrande
85.7

multimodal

BenchmarkDeepSeek-V3GPT-4o
MMMU
69.1

safety

BenchmarkDeepSeek-V3GPT-4o
TruthfulQA
73.5

Category Winners

coding
GPT-4o
38.2
math
GPT-4o
54.5
reasoning
GPT-4o
54.4
general
GPT-4o
70.5
language
GPT-4o
79.9
multimodal
GPT-4o
54.0
safety
GPT-4o
94.3

Compare More Models

Frequently Asked Questions