GPT-4.1 Mini
by OpenAItext+image+file->text3 endpoints
Neura Intelligence Index
17.3/ 100
Rank
#300—
Confidence
highBenchmarks
8(21% coverage)
Overview
GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...
Capabilities
Text generation
Image understanding
Tool use / Function calling
Structured output
Prompt caching
Long context
File input
Top-P sampling
Deterministic seed
Modalities
Input
ImageTextFile
Output
Text
Technical Specifications
Context Window
1.0M tokens
Max Output
32.8K tokens
Knowledge Cutoff
2024-06-30
Tokenizer
GPT
Content Moderation
Enabled
Uptime (24h)
10000.0%
Country
🇺🇸US
Release Date
April 14, 20251y
Supported Parameters
max_completion_tokensmax_tokensresponse_formatseedstructured_outputstemperaturetool_choicetoolstop_p
Benchmark Performance
general
| Benchmark | Score | Source |
|---|---|---|
| MMMLU | 78.5 | verified |
coding
| Benchmark | Score | Source |
|---|---|---|
| SWE-bench Verified | 23.6 | verified |
math
| Benchmark | Score | Source |
|---|---|---|
| AIME 2025 | 40.2 | verified |
reasoning
| Benchmark | Score | Source |
|---|---|---|
| GPQA Diamond | 65 | verified |
| Humanity's Last Exam | 3.7 | verified |
multimodal
| Benchmark | Score | Source |
|---|---|---|
| MMMU | 72.7 | verified |
| CharXiv Reasoning | 56.8 | verified |
agent
| Benchmark | Score | Source |
|---|---|---|
| TAU-Bench Retail | 55.8 | verified |
Performance Over Time
Pricing
Per 1M tokens
Input$0.40
Output$1.60
Cache read$0.10
Added
April 14, 2025
Last synced 9/16/2026