Deepseek V4 Flash is now ~#2 open weight model to Kimi K3 and >50x cheaper

Reddit r/LocalLLaMA Models

Summary

DeepSeek's new V4 Flash model is reportedly the #2 open-weight model behind Kimi K3, offering strong performance at over 50x lower cost ($0.09/$0.18 per 1M tokens) with solid coding and reasoning capabilities.

https://preview.redd.it/h7zv5tb3tmgh1.png?width=2854&format=png&auto=webp&s=507380e8f862c18f10f7c5c84da9e8d1c59139b0 Deepseek's new flash model is unexpectedly cheap and high-performing across useful benchmarks. It's priced at $0.09 / $0.18 per 1M. Truly "intelligence too cheap to meter". Seems to work pretty well on coding, reasoning chat topics for me. How's it holding up for you all in your testing and work?
Original Article

Similar Articles

Qwen3.8-Max matches Kimi K3 and DeepSeek V4 Flash

Reddit r/LocalLLaMA

Qwen3.8-Max, a 2.4T parameter open-weight model, matches Kimi K3 and DeepSeek V4 Flash on benchmarks, excelling in coding and software tasks. Weights release next week, with pricing of $2/M input and $6/M output tokens.