deepseek-ai/DeepSeek-V4-Pro-0813 · Hugging Face
Summary
DeepSeek has released DeepSeek-V4-Pro-0813, a new version of its large language model, available on Hugging Face.
Similar Articles
deepseek-ai/DeepSeek-V4-Flash
DeepSeek releases DeepSeek-V4-Flash and DeepSeek-V4-Pro, new MoE language models supporting 1 million token contexts with improved efficiency and performance.
DeepSeek-V4-Pro-0813-NVFP4 (7 minute read)
NVIDIA releases a quantized version of DeepSeek's V4-Pro-0813 model on Hugging Face, optimized with NVFP4 for efficient deployment in agentic AI and reasoning applications.
deepseek-ai/DeepSeek-V4.1-Flash · Hugging Face
The repository provides prompt encoding and a minimal PyTorch inference implementation for the DeepSeek-V4.1-Flash AI model, including components like vision encoder, MoE, and Hyper-Connections under an MIT License.
deepseek-ai/DeepSeek-V4-Flash-DSpark
DeepSeek releases V4 series of Mixture-of-Experts language models (Pro 1.6T/49B activated, Flash 284B/13B activated) supporting one-million-token context with hybrid attention and speculative decoding, claiming best open-source model performance.
deepseek-ai/DeepSeek-V4-Pro
DeepSeek releases V4-Pro and V4-Flash, Mixture-of-Experts models supporting million-token context with hybrid attention and Muon optimizer.