DeepSeek's new AI model is by far the cheapest of well-known models to run, research firm says (4 minute read)
Summary
DeepSeek's new V4-Flash AI model is reported to be the cheapest well-known model to run, costing 105 times less than Anthropic's Claude Fable 5.
Similar Articles
DeepSeek just popped the American AI bubble.
DeepSeek's V4 Pro model undercuts rivals like GPT-5.5 and Claude Opus by 10-35x on pricing, signaling a deflationary pressure on the AI bubble as margins compress with 'good enough' models at significantly lower cost.
DeepSeek Flash just revolutionized the agent market: 100x cheaper agents
DeepSeek Flash is a new AI model that dramatically reduces the cost of building AI agents by 100x, potentially revolutionizing the agent market.
)
DeepSeek permanently reduced V4 Pro prices by 75%, undercutting leading AI models from OpenAI, Anthropic, and Google, escalating the AI price war.
deepseek-ai/DeepSeek-V4-Flash-0731
DeepSeek released DeepSeek-V4-Flash-0731, a 304B parameter model with enhanced agentic capabilities, priced at $0.14/M input and $0.27/M output, punching above its weight and ranking as the best value-per-intelligence model according to Artificial Analysis.
Is DeepSeek v4 (Flash) really extremely cheap to run? If yes, how?
The user asks why DeepSeek v4 Flash (284B parameters) is so cheap to run compared to smaller models like Qwen 27B, questioning if it's due to pricing dumping or architectural differences. The answer likely involves its MoE architecture and efficient inference techniques.