text-diffusion

Tag

Cards List
#text-diffusion

Exploring More to Solve More: Boosting Diversity in Text Diffusion Models via Entropy-Based Guidance

arXiv cs.CL ↗ · 2026-08-04 Cached

Introduces a training-free Semantic-Aware Kernel Entropy (SAKE) guidance method for text diffusion models, using order-2 Rényi entropy over a kernel Gram matrix to balance fidelity and diversity during sampling. Experiments show improved Pareto frontier and multi-sample performance on reasoning-intensive tasks.

0 favorites 0 likes
#text-diffusion

[Talk] Text Diffusion — Google DeepMind's Brendan O’Donoghue

Reddit r/LocalLLaMA ↗ · 2026-06-11 Cached

DeepMind researcher Brendan O'Donoghue provides an in-depth introduction to text diffusion models, which generate text through iterative denoising. Compared to autoregressive models, they offer lower latency but limited throughput, and demonstrate unique advantages such as self-correction and dynamic computation.

0 favorites 0 likes
#text-diffusion

@omarsar0: This is awesome! I am spending a lot of time on diffusion LLMs these days, so this is perfect timing. I feel like there…

X AI KOLs Following ↗ · 2026-06-10 Cached

Google DeepMind released DiffusionGemma, an open experimental model that generates text in blocks rather than word-by-word, enabling self-correction and faster output.

0 favorites 0 likes
#text-diffusion

The Safety-Aware Denoiser for Text Diffusion Models

arXiv cs.LG ↗ · 2026-05-12 Cached

This paper introduces the Safety-Aware Denoiser (SAD), a framework for integrating safety constraints into text diffusion models during the denoising process. It aims to reduce unsafe generations while preserving quality, addressing a gap in safety research for non-autoregressive models.

0 favorites 0 likes
#text-diffusion

@JulieKallini: Fast Byte Latent Transformer is accepted to ICML 2026! Byte-level LMs promise to free us from subword tokenizers, but d…

X AI KOLs Following ↗ · 2026-05-11 Cached

The Fast Byte Latent Transformer (BLT-D) has been accepted to ICML 2026, introducing a text diffusion method for parallel byte-level decoding to overcome the speed limitations of traditional byte-level language models.

0 favorites 0 likes
← Back to home

Submit Feedback