GPT-4.1 logo

GPT-4.1

Paid
4.5
Inputs: text, code, videoOutputs: text, code
Type
Saas
Company
OpenAI

About GPT-4.1

GPT-4.1 is a family of large language models launched by OpenAI in April 2025, available exclusively through the API. The series includes three models: GPT-4.1, GPT-4.1 mini, and GPT-4.1 nano, offering significant improvements over GPT-4o and GPT-4o mini in coding, instruction following, and long-context comprehension. All models support up to 1 million tokens of context and feature a knowledge cutoff of June 2024. GPT-4.1 achieves state-of-the-art results on benchmarks such as SWE-bench Verified (54.6%), MultiChallenge (38.3%), and Video-MME (72.0% on long, no subtitles). The mini model balances performance with 83% cost reduction and nearly half the latency of GPT-4o, while the nano model is the fastest and cheapest, excelling at classification and autocompletion. The family is optimized for real-world developer tasks, including building reliable AI agents, software engineering, document analysis, and customer request resolution. GPT-4.1 is not available in ChatGPT; its improvements will be gradually incorporated into the ChatGPT version of GPT-4o.

Key Features

Three model sizes: GPT-4.1, GPT-4.1 mini, and GPT-4.1 nano
Up to 1 million token context window
Improved coding performance: 54.6% on SWE-bench Verified
Enhanced instruction following: 38.3% on MultiChallenge
State-of-the-art long-context understanding: 72.0% on Video-MME (long, no subtitles)
Knowledge cutoff of June 2024
Lower latency and reduced cost compared to GPT-4o
Optimized for building AI agents and real-world developer tasks
GPT-4.1 nano is the fastest and cheapest OpenAI model
GPT-4.1 mini beats GPT-4o on many benchmarks with 83% cost reduction

Pros & Cons

Pros
  • Significant coding benchmark improvement (21.4% abs over GPT-4o on SWE-bench)
  • Best-in-class instruction following and long-context comprehension
  • Multiple price/performance points with mini and nano models
  • Lower latency and cost than previous GPT-4 series models
  • 1M token context window enables large document processing
Cons
  • Only available via API, not in ChatGPT
  • GPT-4.5 Preview will be deprecated on July 14, 2025
  • Exact pricing not disclosed in the announcement
  • Nano model has lower accuracy on complex reasoning (9.8% Aider polyglot coding)

Best For

Software engineering and coding tasksBuilding reliable AI agents for complex tasksExtracting insights from large documentsResolving customer requests with minimal hand-holdingClassification and autocompletion (especially GPT-4.1 nano)Instruction following in multi-step workflowsLong-context reasoning with multimodal inputs (video, text)

Alternatives to GPT-4.1

FAQ

Is GPT-4.1 available in ChatGPT?
No, GPT-4.1 is only available via the API. Improvements in instruction following and coding are gradually being incorporated into the ChatGPT version of GPT-4o.
What are the different GPT-4.1 models?
There are three models: GPT-4.1 (flagship), GPT-4.1 mini (balanced performance and cost), and GPT-4.1 nano (fastest and cheapest).
What is the context window size for GPT-4.1?
All GPT-4.1 models support up to 1 million tokens of context.
When was GPT-4.1 released?
GPT-4.1 was announced on April 14, 2025.
How does GPT-4.1 compare to GPT-4o in coding?
GPT-4.1 scores 54.6% on SWE-bench Verified, an improvement of 21.4% absolute over GPT-4o.
What will happen to GPT-4.5 Preview?
GPT-4.5 Preview will be deprecated on July 14, 2025, as GPT-4.1 offers similar or better performance at lower cost and latency.