DFlash 2 available for Qwen 3.8 27B and Muse Glimmer
Summary
DFlash 2, a model optimization tool by z-lab, is now available for Qwen 3.8 27B and Muse Glimmer models, enhancing text generation capabilities.
View Cached Full Text
Cached at: 08/18/26, 10:35 PM
DFlash 2 - a z-lab Collection
Source: https://huggingface.co/collections/z-lab/dflash-2
![]()
z-lab’s Collections
updatedabout 1 hour ago
Keep Drafting Parallel
- —
#### z-lab/Qwen3.8-27B-DFlash2 Text Generation• 2B• Updatedabout 2 hours ago • 150 • 11 - —
#### z-lab/Muse-Glimmer-30B-DFlash2 Text Generation• 3B• Updatedabout 2 hours ago • 244 • 2 - —
#### z-lab/Qwen3.8-27B-DFlash2-GGUF Text Generation• 2B• Updatedabout 1 hour ago • 6 - —
#### z-lab/Muse-Glimmer-30B-DFlash2-GGUF Text Generation• 3B• Updatedabout 1 hour ago
Similar Articles
DFlash makes Qwen3.6 27B 2.2x faster with no quality loss
DFlash is a method that accelerates Qwen3.6 27B model inference by 2.2x without quality degradation.
z-lab/Qwen3.6-27B-DFlash
This article introduces Qwen3.6-27B-DFlash, a specialized drafter model for DFlash, a novel speculative decoding method using block diffusion to accelerate inference speed. It provides installation instructions for vLLM and SGLang to enable parallel drafting with the target Qwen3.6-27B model.
incoai/Qwen3.8-27B-DFlash2
DFlash 2 is a block-diffusion draft model for speculative decoding that improves inference speed for the Qwen3.8-27B language model, offering higher acceptance length and throughput in benchmarks.
z-lab/Qwen3.8-27B-DFlash2
Introduces DFlash 2, a block-diffusion drafter for speculative decoding with the Qwen3.8-27B model, demonstrating improved acceptance length and throughput in benchmarks.
@zhijianliu_: DFlash for Qwen3.6-35B-A3B just dropped The community was running the day-1 preview before we even finished training. N…
Z-lab releases DFlash for Qwen3.6-35B-A3B, a model fine-tuning/compression technique, with training complete and weights now available on GitHub and HuggingFace.