AI Models

DeepSeek Makes Permanent 75% Discount on V4-Pro AI Model

Chinese AI startup DeepSeek announced it will permanently apply a 75% discount on its flagship V4-Pro model, a price cut originally scheduled to expire at the end of May. The move is expected to intensify competition in the global AI market as Chinese firms aggressively court developers with lower pricing.

Neura News

Neura News

Neura Market Editorial

May 24, 20262 min read

Originally reported by bloomberg.com

DeepSeek Makes Permanent 75% Discount on V4-Pro AI Model

Chinese artificial intelligence startup DeepSeek has confirmed it will permanently maintain a 75% discount on its flagship V4-Pro model, a pricing cut that was originally set to expire at the end of May. The decision, posted on DeepSeek's website, locks in access to the model at one quarter of its original price for developers.

DeepSeek's Pricing Strategy

The V4-Pro is DeepSeek's most advanced large language model, designed for a wide range of generative AI tasks including text generation, code completion, and reasoning. The model competes directly with offerings from OpenAI, Anthropic, Google, and other global players.

Instead, the company opted to make the price cut permanent. The move is likely aimed at building long-term loyalty among developers and startups who are sensitive to API pricing.

Competitive Implications

The #1 Newsletter in AI

Stay ahead of the AI curve

The most important updates, news, and content — delivered weekly.

No spam. Unsubscribe anytime.

DeepSeek's decision comes amid a broader price war in the AI industry, particularly between Chinese companies and their Western counterparts. Chinese AI startups like DeepSeek, Alibaba's Tongyi Qianwen, and Baidu's Ernie have been undercutting the pricing of international models in an effort to gain market share.

Background on DeepSeek

DeepSeek is a Beijing-based AI startup that has quickly emerged as one of the most prominent Chinese players in large language models. The company was founded in 2023 by alumni from major Chinese tech firms and has raised significant venture capital funding. Its V4-Pro model has been benchmarked against leading models from the US and has performed competitively on several industry tests.

As the AI industry continues to evolve, pricing decisions like DeepSeek's will likely shape the competitive dynamics between Chinese and global players. For now, developers can expect to pay a quarter of the original price for one of the more capable models on the market.

Related on Neura Market

  • AI Models Directory, Browse and compare leading language models from providers worldwide.
  • Pricing & Deals, Track discounts and pricing updates across AI tools and platforms.

More from Neura News

AI Models

Google Unveils Gemini 3.6 Flash, 3.5 Flash-Lite, and Cyber Model

Google has released three new Gemini models: 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber. The 3.6 Flash model offers improved coding and knowledge work with 17% fewer output tokens and lower costs. The 3.5 Flash-Lite is the fastest in the series at 350 tokens per second, designed for high-throughput agentic tasks. The 3.5 Flash Cyber model, available only to governments and trusted partners via CodeMender, focuses on finding and fixing cybersecurity vulnerabilities. Google also noted that Gemini 3.5 Pro is being tested with partners and that pre-training for Gemini 4 has begun.

Jul 21·5 min read
AI Models

Alibaba Qwen-Image-3.0 renders infographics and tiny text in one pass

Alibaba's Qwen team released Qwen-Image-3.0, an image generator designed for practical applications like newspaper layouts and complex infographics. The model processes prompts of up to 4,500 tokens and can render legible text as small as ten pixels, mathematical formulas, and twelve languages in a single pass. It is currently available through invite-only API access, with plans to integrate it into first-party apps like Qwen Chat soon.

Jul 21·4 min read
AI Models

Google Unveils Gemini 3.6 Flash, 3.5 Flash-Lite, and Cyber Model

Google DeepMind has introduced three new Gemini models: 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber. The 3.6 Flash model offers improved coding and multimodal performance with 17% fewer output tokens and lower cost. The 3.5 Flash-Lite is the fastest in its series at 350 output tokens per second, designed for high-throughput agentic tasks. The 3.5 Flash Cyber, fine-tuned for cybersecurity, will be available exclusively to governments and trusted partners via the CodeMender agent.

Jul 21·6 min read