DeepSeek V4 Models Close Gap to Frontier AI
Chinese AI laboratory DeepSeek unveiled preview editions of its latest large language model, DeepSeek V4. This includes two variants: V4 Flash and V4 Pro. The release updates last year's V3.2 model and the R1 reasoning model, both of which gained widespread attention in the AI community.
Model Architecture and Scale
DeepSeek describes V4 Flash and V4 Pro as mixture-of-experts systems. Each supports a context window of 1 million tokens. That capacity handles extensive codebases or lengthy documents in single prompts. The mixture-of-experts design activates a subset of parameters for each task. This reduces costs during inference.
The V4 Pro holds 1.6 trillion total parameters, with 49 billion active. It stands as the largest open-weight model released so far. It surpasses Moonshot AI's Kimi K 2.6 at 1.1 trillion parameters, MiniMax's M1 with 456 billion, and exceeds DeepSeek V3.2's 671 billion by more than double. V4 Flash, the compact option, contains 284 billion parameters, 13 billion active.
DeepSeek, founded as a key player in China's AI efforts, focuses on open-source releases to compete globally. Past models like V3.2 demonstrated strong performance at low costs, drawing developers and researchers.
Performance on Benchmarks
According to DeepSeek, the new models outperform V3.2 in efficiency and results thanks to design upgrades. They have nearly matched leading open and closed models on reasoning tests. The V4-Pro-Max variant beats other open-source competitors across reasoning evaluations. It also tops OpenAI's GPT-5.2 and Gemini 3.0 Pro in certain tasks.
On coding competition benchmarks, both V4 models deliver results similar to GPT-5.4. However, they trail slightly on knowledge assessments against GPT-5.4 and Google's Gemini 3.1 Pro. DeepSeek notes this indicates a development path about 3 to 6 months behind top frontier models.
Stay ahead of the AI curve
The most important updates, news, and content — delivered weekly.
No spam. Unsubscribe anytime.
Unlike many proprietary rivals, V4 Flash and V4 Pro handle text exclusively. They lack features for audio, video, or image processing found in some closed systems.
Pricing Advantages
DeepSeek V4 offers costs far below current frontier options. V4 Flash charges $0.14 per million input tokens and $0.28 per million output tokens. Those rates beat GPT-5.4 Nano, Gemini 3.1 Flash, GPT-5.4 Mini, and Claude Haiku 4.5.
V4 Pro prices at $0.145 per million input tokens and $3.48 per million output tokens. It undercuts Gemini 3.1 Pro, GPT-5.5, Claude Opus 4.7, and GPT-5.4. Such affordability stems from the efficient architecture and open-weight approach, making high capability accessible.
Broader Context
The announcement arrived one day after U.S. authorities charged China with large-scale theft of American AI intellectual property via thousands of proxy accounts. Separately, Anthropic and OpenAI have accused DeepSeek of distilling, a process that copies elements from their models.
DeepSeek continues to push boundaries in open AI development from China. Its models provide alternatives to dominant U.S. providers, emphasizing scale, cost, and benchmark parity.

