Unprecedented Cost Efficiency
At the absolute operational center of this release is V4-Flash’s highly disruptive pricing structure. According to the San Francisco-based research firm Artificial Analysis, the model costs an average of just 3 cents to complete a full benchmark test battery. This represents a massive cost reduction compared to its rivals, operating over 100 times cheaper than Anthropic’s Claude Fable 5 ($3.15 per test) and significantly undercutting OpenAI’s GPT-5.6 Sol ($1.86) and Moonshot AI’s Kimi K3 (86 cents). The baseline token pricing is set aggressively low at $0.14 per million input tokens and $0.28 per million output tokens.
Performance Benchmarks and Capabilities
While V4-Flash dominates in cost efficiency, benchmark data indicates it operates just below the absolute frontier of AI capabilities. The model scored a 50 out of 100 on the Intelligence Index—a battery evaluating coding, reasoning, and workplace tasks—placing it exactly on par with Google’s Gemini 3.6 Flash. However, it trails more expensive, high-tier models like OpenAI’s GPT-5.6 and Anthropic’s Claude Opus 5 by roughly nine points, reinforcing its position as a high-speed, cost-effective solution for everyday volume tasks rather than complex reasoning.
Market Strategy and Future Rollouts
The introduction of V4-Flash deliberately intensifies an ongoing global AI price war that DeepSeek originally triggered following the massive success of its R1 model in early 2025. By leveraging highly efficient training and inference engineering, the startup aims to capture massive enterprise volume—such as chatbots and back-office automation—where fractions of a cent compound rapidly. Looking ahead, the company has confirmed the development of a heavier, more capable “V4-Pro” model designed to tackle complex reasoning, though an official release date remains unannounced.







