Antirez Deepseek 4.1 flash gguf on HF
Summary
Antirez has uploaded the quantized gguf version of Deepseek 4.1 flash model to Hugging Face, with partial availability and questions on usage.
Similar Articles
Deepseek V4 Flash 2, 3 and 4 bits GGUFs
GGUF quantizations of DeepSeek V4 Flash in 2-bit, 3-bit, and 4-bit precisions, made available on Hugging Face for local inference with tools like llama.cpp and Ollama.
antirez/deepseek-v4-gguf
Antirez released GGUF quantizations of DeepSeek V4 Flash specifically tailored for the DS4 inference engine, providing optimized configurations for different RAM sizes and enabling local execution of the large MoE model.
Bartowski has delivered DS4 GGUF
Bartowski has released a GGUF quantized version of DeepSeek-V4-Flash, inviting comparison with Antirez's version.
unsloth/DeepSeek-V4-Flash-0731-GGUF
Unsloth teases the upcoming release of DeepSeek V4 Flash GGUF quantized model on Hugging Face.
huihui-ai/Huihui-DeepSeek-V4-Flash-0731-abliterated-GGUF
A model card for Huihui-DeepSeek-V4-Flash-0731-abliterated-GGUF, an abliterated (uncensored) GGUF quantized variant of DeepSeek-V4-Flash, designed for local use with llama.cpp and ds4.