Crimson AI NewsA CrimsonLingua Network service
EN ع
← Back to news
DeepSeek

DeepSeek V4 Launches with 1.6T Parameters, Claims 10-50x Cheaper Pricing

AI By Crimson AI DeepSeek Blog 10 August 2026 · 00:32 20 views
Share: X Telegram

DeepSeek unveiled its V4 model family on April 24, 2026, featuring two text-only variants with up to 1.6 trillion parameters and a 1M-token context window, priced significantly below Western rivals.

DeepSeek V4 Launches with 1.6T Parameters, Claims 10-50x Cheaper Pricing

Key points

DeepSeek has officially launched its V4 model family, marking a significant push in the AI efficiency race. The lineup includes two variants: V4-Pro with 1.6 trillion total parameters (49 billion active per token) and V4-Flash with 284 billion total parameters (13 billion active). Both are text-only and support a 1 million token context window, positioning them as direct competitors to models like OpenAI's GPT-5.4 and Anthropic's Claude Opus 4.8.

The architecture introduces two key innovations: Manifold-Constrained Hyper-Connections (mHC), which stabilizes training of massive MoE models by limiting signal amplification to under 2x (versus an unconstrained 3000x), and DeepSeek Sparse Attention (DSA), which compresses context into coarse summaries and applies full attention only to relevant regions, enabling near-linear cost scaling for long contexts.

Pricing is a major differentiator. V4-Pro costs $0.435 per million input tokens (cache miss) and $0.87 per million output tokens, while V4-Flash is $0.14 and $0.28 respectively. This is 5.7x to 28.7x cheaper than GPT-5.4 and Claude Opus 4.8 on input and output. The models are open-sourced under the MIT license, and DeepSeek claims they can run locally on dual RTX 4090s or a single RTX 5090, though the full 1.6T model remains a data-center workload.

DeepSeek has also optimized V4 for domestic Chinese silicon, including Huawei Ascend 950PR and Cambricon MLU chips, due to US export restrictions on advanced Nvidia GPUs. The Ascend 950PR reportedly delivers 2.87x the compute performance of the Nvidia H20. This marks a strategic shift toward AI semiconductor independence in China.

Benchmark figures cited in the source are based on leaked internal data and await third-party verification. The company also retired legacy model IDs 'deepseek-chat' and 'deepseek-reasoner' on July 24, 2026.

ModelTotal ParamsActive per TokenContext WindowInput Price (per 1M)Output Price (per 1M)
V4-Pro1.6T49B1M tokens$0.435$0.87
V4-Flash284B13B1M tokens$0.14$0.28
GPT-5.4N/AN/A270K tokens$2.50$15.00
Claude Opus 4.8N/AN/AN/A$5.00$25.00
Source
DeepSeek · DeepSeek Blog
Related news
DeepSeek
DeepSeek 4 Aug 2026

DeepSeek Retires Legacy API Aliases: Price Changes and Migration Guide

DeepSeek retired the legacy API aliases 'deepseek-chat' and 'deepseek-reasoner' on July 24, 2026, at 15:59 UTC. Developers must up...

33
DeepSeek
DeepSeek 2 Aug 2026

DeepSeek-V4-Flash Goes Official: Agent Benchmarks Surpass V4-Pro-Preview

DeepSeek has launched the official deepseek-v4-flash API in public beta, featuring improved agent capabilities that outperform V4-...

33
DeepSeek
DeepSeek 27 Jul 2026

DeepSeek vs Stepfun: Two Chinese AI Labs, Two Divergent Strategies for 2026

DeepSeek focuses on reasoning and low-cost APIs, while Stepfun bets on multimodal generation including video and audio. Here's how...

29