All Comparisons

GPT-4.1 vs GPT-4o mini

GPT-4.1
OpenAI πŸ‡ΊπŸ‡Έ
32.8
#238β€”
high
$2.00/M in Β· $8.00/M out
1.0M context
GPT-4o mini
OpenAI πŸ‡ΊπŸ‡Έ
41.5
#200β€”
high
$0.15/M in Β· $0.60/M out
128K context
Most Affordable
GPT-4o mini
$0.15/M input
Largest Context
GPT-4.1
1.0M tokens
Best Benchmark Score
GPT-4o mini
41.5/100

Model Specifications

SpecGPT-4.1GPT-4o mini
ProviderOpenAIOpenAI
Release DateApr 2025Jul 2024
Knowledge Cutoff2024-06-01Oct 2023
ParametersUndisclosedunknown
Context Window1.0M128K
Max Outputβ€”β€”
Open SourceNoNo
Licenseproprietaryproprietary
Tokenizerβ€”β€”
Modalityβ€”β€”
Reasoning ModelNoNo
ModeratedNoNo

Pricing Comparison

Price (per 1M tokens)GPT-4.1GPT-4o mini
Input$2.00$0.15
Output$8.00$0.60

API Provider Pricing

Capabilities

FeatureGPT-4.1GPT-4o mini
Text Input
Image Input (Vision)
Audio Input
Video Input
File Input
Image Output
Audio Output
Tool Use
Structured Output (JSON)
Streaming
Reasoning Tokens
Web Search
Temperature Control
Top-P Sampling
Stop Sequences
Seed (Reproducibility)
Log Probabilities

Category Comparison

Benchmark-by-Benchmark

general

BenchmarkGPT-4.1GPT-4o mini
MMMLU
87.3
β€”
IFEvalβ€”
80.1
MMLU-Proβ€”
63.2
Chatbot Arena Eloβ€”
1219

coding

BenchmarkGPT-4.1GPT-4o mini
SWE-bench Verified
54.6
β€”
HumanEvalβ€”
87.2

math

BenchmarkGPT-4.1GPT-4o mini
AIME 2025
46.4
β€”
MATHβ€”
70.2
GSM8Kβ€”
93.2

reasoning

BenchmarkGPT-4.1GPT-4o mini
GPQA Diamond
66.3
40.2
Humanity's Last Exam
5.4
β€”
ARC-Challengeβ€”
93.1
BigBench-Hardβ€”
78.5

multimodal

BenchmarkGPT-4.1GPT-4o mini
MMMU
74.8
β€”
CharXiv Reasoning
56.7
β€”

agent

BenchmarkGPT-4.1GPT-4o mini
TAU-Bench Retail
68
β€”

language

BenchmarkGPT-4.1GPT-4o mini
WinoGrandeβ€”
82.3
HellaSwagβ€”
91.2

safety

BenchmarkGPT-4.1GPT-4o mini
TruthfulQAβ€”
65.8

Category Winners

coding
GPT-4o mini
61.6
math
GPT-4o mini
48.4
reasoning
GPT-4.1
27.5
general
GPT-4.1
66.0
language
GPT-4o mini
35.0
multimodal
GPT-4.1
40.2
safety
GPT-4o mini
62.9
agent
GPT-4.1
50.7

Compare More Models

Frequently Asked Questions