Deepseek v4.1 flash finally has engrams, what do you expect from 4.1 pro?

Reddit r/LocalLLaMA Models

Summary

Speculation on Deepseek v4.1 pro's specifications and future developments, including parameter counts and engram technology.

Maybe 2.6T -2.8T params including 1-1.1 T engrams and fable 5.0 level performance? Maybe v4.2 or 4.5 will have engram gradient modification?
Original Article

Similar Articles

DeepSeek V4 Pro 0813 quietly released

Hacker News Top

DeepSeek has quietly released V4 Pro (0813), with API documentation showing support for the Responses API format and integration with Codex, enabling use of deepseek-v4-flash/pro models.

DeepSeek v4.1 Flash

Hacker News Top

DeepSeek has introduced DeepSeek-V4.1-Flash, a new AI model designed for enhanced capability, faster inference, native visual understanding, and scalability as part of their latest architecture family.