BotBeat
...
← Back

> ▌

DeepSeekDeepSeek
PRODUCT LAUNCHDeepSeek2026-05-01

DeepSeek Releases V4 Models with 1M-Token Context and Aggressive Pricing Strategy

Key Takeaways

  • ▸DeepSeek V4 models feature 1 million token context and come in two variants: Pro (1.6T/49B active) and Flash (284B/13B active)
  • ▸Pricing significantly undercuts competitors, with Flash at $0.14/$0.28 per million tokens and Pro at $1.74/$3.48
  • ▸Both models achieve dramatic efficiency improvements—27-10% of V3.2's FLOPs and KV cache size in million-token contexts
Source:
Hacker Newshttps://simonw.substack.com/p/deepseek-v4-and-the-end-of-the-openaimicrosoft↗

Summary

Chinese AI lab DeepSeek has released two preview models from its anticipated V4 series: DeepSeek-V4-Pro and DeepSeek-V4-Flash. Both feature 1 million token context windows and are built on Mixture of Experts architecture, with Pro containing 1.6 trillion total parameters (49B active) and Flash containing 284 billion total (13B active). Both models are released under the MIT license as open weights.

The models demonstrate significant efficiency improvements over V3.2, with DeepSeek-V4-Pro achieving only 27% of the single-token FLOPs and 10% of the KV cache size of its predecessor in 1M-token contexts. Most notably, the pricing is substantially lower than competitors: Flash costs $0.14-$0.28 per million tokens, making it cheaper than OpenAI's GPT-5.4 Nano, while Pro at $1.74-$3.48 per million tokens is the least expensive frontier-class model available.

According to DeepSeek's benchmarks, V4-Pro is competitive with leading models from OpenAI, Google, and Anthropic, though DeepSeek acknowledges a 3-6 month development gap from state-of-the-art. With V4-Pro being the largest open weights model to date and both models using MIT licensing, the release represents a significant democratization of access to capable large language models.

  • Released under MIT license as open weights, democratizing access to frontier-class model capabilities
  • V4-Pro performance is competitive with leading models despite being 3-6 months behind state-of-the-art

Editorial Opinion

DeepSeek V4 represents a watershed moment in the LLM market: frontier-class performance at commodity pricing, with open weights licensing. The dramatic cost advantage and efficiency gains could force significant price competition across the industry and accelerate deployment of capable AI systems globally. While the acknowledged developmental gap from pure frontier models suggests room for continued improvement, DeepSeek has effectively decoupled price from capability—a shift that threatens the incumbent business models of OpenAI, Google, and Anthropic.

Large Language Models (LLMs)Generative AIMarket TrendsProduct LaunchOpen Source

More from DeepSeek

DeepSeekDeepSeek
RESEARCH

Researchers Discover DeepSeek-Powered Autonomous Cyberattack Campaign

2026-08-01
DeepSeekDeepSeek
RESEARCH

DeepSeek V4 Flash Achieves Parity with GPT-5.6 on Agentic Memory Benchmark at 20x Lower Cost

2026-07-31
DeepSeekDeepSeek
UPDATE

DeepSeek Releases V4-Flash: Optimized LLM for Speed and Efficiency

2026-07-31

Comments

Suggested

Hugging FaceHugging Face
OPEN SOURCE

Strangers Pretrain 15M-Parameter Language Model Using GitHub Actions and Hugging Face PRs

2026-08-02
General AI ResearchGeneral AI Research
RESEARCH

Research Identifies Fundamental Trilemma: LLM Safeguards Cannot Simultaneously Provide Reliable Safety, Useful Capability, and Open Access

2026-08-02
Alibaba (Cloud)Alibaba (Cloud)
INDUSTRY REPORT

Token Diplomacy: China Positions Open-Source AI as Global Strategic Resource

2026-08-02
← Back to news
© 2026 BotBeat
AboutPrivacy PolicyTerms of ServiceContact Us