Deepseek V4 Flash is now ~#2 open weight model to Kimi K3 and >50x cheaper
Summary
DeepSeek's new V4 Flash model is reportedly the #2 open-weight model behind Kimi K3, offering strong performance at over 50x lower cost ($0.09/$0.18 per 1M tokens) with solid coding and reasoning capabilities.
Similar Articles
We Tested DeepSeek V4 Pro and Flash Against Claude Opus 4.7 and Kimi K2.6 (11 minute read)
DeepSeek released V4 Pro and V4 Flash under MIT license on April 24, 2026. In benchmarks against Claude Opus 4.7 and Kimi K2.6, V4 Pro scored 77/100 at $2.25, placing between Opus 4.7 (91) and Kimi K2.6 (68), while V4 Flash scored 60/100 at $0.02, the cheapest in the comparison, with a 75% discount on V4 Pro through May 31.
Qwen3.8-Max matches Kimi K3 and DeepSeek V4 Flash
Qwen3.8-Max, a 2.4T parameter open-weight model, matches Kimi K3 and DeepSeek V4 Flash on benchmarks, excelling in coding and software tasks. Weights release next week, with pricing of $2/M input and $6/M output tokens.
@TheAhmadOsman: DeepSeek V4 Flash 0731 beats Qwen 3.8 27B btw
Ahmad tweets that DeepSeek V4 Flash 0731 outperforms Qwen 3.8 27B, and lists other models like Kimi K3, GLM 5.2, and MiniMax H3.
DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
DeepSeek plans to officially release the V4.1 Flash model around September 10, 2026, which surpasses V4 Pro in performance, cost, and speed, with adjusted pricing for off-peak and peak hours.
Deepseek V4.1 Flash Release Video [Made with Deepseek V4.1 Flash]
The author benchmarks Deepseek V4.1 Flash on motion video generation, finding it has improved to nearly match Opus class models compared to earlier versions like Kimi K3.