GPT-4.1
by OpenAItext+image+file->text3 endpoints
Neura Intelligence Index
32.5/ 100
Rank
#256—
Confidence
highBenchmarks
8(21% coverage)
Overview
GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...
Capabilities
Text generation
Image understanding
Tool use / Function calling
Structured output
Prompt caching
Long context
File input
Top-P sampling
Deterministic seed
Modalities
Input
ImageTextFile
Output
Text
Technical Specifications
Context Window
1.0M tokens
Max Output
32.8K tokens
Knowledge Cutoff
2024-06-30
Tokenizer
GPT
Content Moderation
Enabled
Uptime (24h)
10000.0%
Country
🇺🇸US
Release Date
April 14, 20251y
Supported Parameters
max_completion_tokensmax_tokensresponse_formatseedstructured_outputstemperaturetool_choicetoolstop_p
Benchmark Performance
general
| Benchmark | Score | Source |
|---|---|---|
| MMMLU | 87.3 | verified |
coding
| Benchmark | Score | Source |
|---|---|---|
| SWE-bench Verified | 54.6 | verified |
math
| Benchmark | Score | Source |
|---|---|---|
| AIME 2025 | 46.4 | verified |
reasoning
| Benchmark | Score | Source |
|---|---|---|
| GPQA Diamond | 66.3 | verified |
| Humanity's Last Exam | 5.4 | verified |
multimodal
| Benchmark | Score | Source |
|---|---|---|
| MMMU | 74.8 | verified |
| CharXiv Reasoning | 56.7 | verified |
agent
| Benchmark | Score | Source |
|---|---|---|
| TAU-Bench Retail | 68 | verified |
Performance Over Time
Pricing
Per 1M tokens
Input$2.00
Output$8.00
Cache read$0.50
Added
April 14, 2025
Last synced 9/20/2026