DeepSeek V4 Flash 0731 in Hermes Agent and one prompt, took 32 minutes and cost 0.07$, this model is so cheap to the point where 2 dollars can last you a full day.

Reddit r/singularity Models

Summary

DeepSeek V4 Flash 0731 is highlighted for its extremely low cost when used in Hermes Agent, completing a prompt in 32 minutes for just $0.07, making it highly affordable for sustained use.

No content available
Original Article

Similar Articles

DeepSeek-V4-Flash-0731

Product Hunt

DeepSeek announces DeepSeek-V4-Flash-0731, a frontier agent intelligence model positioned as offering advanced capabilities at Flash-level pricing.

deepseek-ai/DeepSeek-V4-Flash-0731

Simon Willison's Blog

DeepSeek released DeepSeek-V4-Flash-0731, a 304B parameter model with enhanced agentic capabilities, priced at $0.14/M input and $0.27/M output, punching above its weight and ranking as the best value-per-intelligence model according to Artificial Analysis.

Is DeepSeek v4 (Flash) really extremely cheap to run? If yes, how?

Reddit r/LocalLLaMA

The user asks why DeepSeek v4 Flash (284B parameters) is so cheap to run compared to smaller models like Qwen 27B, questioning if it's due to pricing dumping or architectural differences. The answer likely involves its MoE architecture and efficient inference techniques.