DeepSeek V4 Launches with Flash and Pro Versions, Pricing Unveiled
A major update has arrived in the large model space. Leading Chinese AI company DeepSeek today launched its new flagship model DeepSeek V4. The key highlight is its differentiated strategy, which addresses diverse needs—from high-frequency lightweight tasks to complex reasoning—through two versions: Flash and Pro. With aggressive pricing, it once again sets a new benchmark for AI commercial costs.
Model Matrix: Differentiated Positioning of Flash and Pro
DeepSeek V4 integrates and upgrades the original deepseek-chat and deepseek-reasoner models, officially splitting them into two versions:
DeepSeek-V4-Flash: Focused on extreme cost-efficiency and high throughput, ideal for fast-response general conversations and basic text tasks.
DeepSeek-V4-Pro: Optimized for complex logic, deep reasoning, and high-performance computing, offering stronger reasoning and processing capabilities.
Both models support thinking mode (except in specific scenarios), JSON output, tool calls, dialogue prefix continuation (Beta), and a context window of up to 1M tokens with a maximum output of 384K tokens, providing a solid foundation for complex engineering implementations.

Pricing System: Transparent, Tiered Billing
DeepSeek’s pricing is clear and significantly reduces the marginal cost for long-term enterprise API calls through a caching mechanism. Here are the billing rates per million tokens (in RMB):
ModelInput (Cache Hit)Input (Cache Miss)OutputDeepSeek-V4-Flash0.2 RMB1 RMB2 RMBDeepSeek-V4-Pro1 RMB12 RMB24 RMBNote: Fees are deducted first from the free balance, then from the paid balance.
Industry Analysis: Why This Price Is a Milestone
The pricing logic shows DeepSeek is encouraging developers to reduce compute waste by optimizing cache usage, enabling fine-grained cost control through the significant price gap between cached and uncached inputs.
For developers, the Flash version at 1 RMB per million tokens (cache miss) dramatically lowers the barrier to accessing top-tier MoE model architectures. Meanwhile, the Pro version’s pricing for complex reasoning tasks provides a domestic solution for building knowledge-base Q&A systems, autonomous agents, and other advanced applications, balancing performance and cost.
Important Compatibility Note
The official reminder: the original model names deepseek-chat and deepseek-reasoner will be formally deprecated. To ensure a smooth transition, developers can now call deepseek-v4-flash (corresponding to non-thinking mode) and deepseek-v4-pro (corresponding to thinking mode).
For detailed API information and migration guides, developers can visit the official DeepSeek API documentation for more details.
Related article
Suno to Watermark Songs Amid Legal Battles
Suno, the platform enabling users to generate AI-created music, has unveiled new features to label platform-produced tracks, restrict downloads, and update community standards to curb unauthorized replicas. These updates arrive as Suno confronts mult
Musk Admits Grok Build Leaked User Code, Promises to Erase All Historical Data
Elon Musk directly addressed the privacy controversy surrounding Grok Build, beginning with a simple "True" to confirm the incident's validity. He pledged that all user data previously uploaded to SpaceXAI would be permanently erased, stating, "not a
U.S. Stocks Hit Historic Milestone as AI and Aerospace Giants Prepare for Trillion-Dollar Debut
Elon Musk, Sam Altman, and Dario Amodei, three titans of the technology sector, are advancing toward initial public offerings for their respective ventures. With SpaceX, OpenAI, and Anthropic—three industry behemoths nearing trillion-dollar valuation
Related Special Topic Recommendations
Comments (0)
0/500
A major update has arrived in the large model space. Leading Chinese AI company DeepSeek today launched its new flagship model DeepSeek V4. The key highlight is its differentiated strategy, which addresses diverse needs—from high-frequency lightweight tasks to complex reasoning—through two versions: Flash and Pro. With aggressive pricing, it once again sets a new benchmark for AI commercial costs.
Model Matrix: Differentiated Positioning of Flash and Pro
DeepSeek V4 integrates and upgrades the original deepseek-chat and deepseek-reasoner models, officially splitting them into two versions:
DeepSeek-V4-Flash: Focused on extreme cost-efficiency and high throughput, ideal for fast-response general conversations and basic text tasks.
DeepSeek-V4-Pro: Optimized for complex logic, deep reasoning, and high-performance computing, offering stronger reasoning and processing capabilities.
Both models support thinking mode (except in specific scenarios), JSON output, tool calls, dialogue prefix continuation (Beta), and a context window of up to 1M tokens with a maximum output of 384K tokens, providing a solid foundation for complex engineering implementations.

Pricing System: Transparent, Tiered Billing
DeepSeek’s pricing is clear and significantly reduces the marginal cost for long-term enterprise API calls through a caching mechanism. Here are the billing rates per million tokens (in RMB):
ModelInput (Cache Hit)Input (Cache Miss)OutputDeepSeek-V4-Flash0.2 RMB1 RMB2 RMBDeepSeek-V4-Pro1 RMB12 RMB24 RMBNote: Fees are deducted first from the free balance, then from the paid balance.
Industry Analysis: Why This Price Is a Milestone
The pricing logic shows DeepSeek is encouraging developers to reduce compute waste by optimizing cache usage, enabling fine-grained cost control through the significant price gap between cached and uncached inputs.
For developers, the Flash version at 1 RMB per million tokens (cache miss) dramatically lowers the barrier to accessing top-tier MoE model architectures. Meanwhile, the Pro version’s pricing for complex reasoning tasks provides a domestic solution for building knowledge-base Q&A systems, autonomous agents, and other advanced applications, balancing performance and cost.
Important Compatibility Note
The official reminder: the original model names deepseek-chat and deepseek-reasoner will be formally deprecated. To ensure a smooth transition, developers can now call deepseek-v4-flash (corresponding to non-thinking mode) and deepseek-v4-pro (corresponding to thinking mode).
For detailed API information and migration guides, developers can visit
Suno to Watermark Songs Amid Legal Battles
Suno, the platform enabling users to generate AI-created music, has unveiled new features to label platform-produced tracks, restrict downloads, and update community standards to curb unauthorized replicas. These updates arrive as Suno confronts mult
Musk Admits Grok Build Leaked User Code, Promises to Erase All Historical Data
Elon Musk directly addressed the privacy controversy surrounding Grok Build, beginning with a simple "True" to confirm the incident's validity. He pledged that all user data previously uploaded to SpaceXAI would be permanently erased, stating, "not a
U.S. Stocks Hit Historic Milestone as AI and Aerospace Giants Prepare for Trillion-Dollar Debut
Elon Musk, Sam Altman, and Dario Amodei, three titans of the technology sector, are advancing toward initial public offerings for their respective ventures. With SpaceX, OpenAI, and Anthropic—three industry behemoths nearing trillion-dollar valuation





Home






