Why there isn't any top LLM providers investing on diffusion LLM?

Reddit r/singularity News

Summary

This article questions why major LLM providers are not investing in Diffusion LLMs despite recent advancements like Mercury 2. It explores potential fundamental issues or hardware bottlenecks hindering broader adoption.

A year ago, I would’ve said Diffusion LLMs were an interesting idea but still far from practical. They’re still pretty rough, but Mercury 2 now makes it seem like they might finally be getting close to usable. That said, aside from Meta, Ant, and Inception/Mercury, it doesn’t seem like many labs are seriously investing in them — especially the major ones like OpenAI, Anthropic, Google, xAI, or even architecture-focused teams like DeepSeek and Kimi. I’m not very familiar with DLLMs, so I’m curious: why is that? Are there still fundamental issues with the paradigm that make them unlikely to become even second-tier models? Or is current hardware stack a bottleneck for DLLMs training/inference? Or are other labs just working on it quietly and not there yet?
Original Article

Similar Articles

Why current LLM costs are not sustainable

Hacker News Top

The article argues that current high LLM pricing is unsustainable due to diminishing performance gains, the rise of open-weight models, specialized AI chips reducing inference costs, and zero switching costs, predicting significant price drops as competition intensifies.

Who is leading the pack for best llm gateway 2k26?

Reddit r/LocalLLaMA

The article explores the search for the best LLM gateways in 2026, emphasizing enterprise-grade features like prompt caching, granular cost attribution, and support for reasoning models.

These startups are chasing the next big thing in LLMs

MIT Technology Review

MIT Technology Review reports on a wave of startups pursuing post-transformer architectures for LLMs, as the dominant model family faces growing costs, energy use, and context-length limits. Companies like Subquadratic aim to build the next generation of AI.

LLMs are stuck in a groupthink groove. This startup is trying to get them out.

MIT Technology Review

A startup called Springboards has developed an LLM named Flint that aims to produce more varied and creative responses than mainstream models, addressing a widespread issue of homogeneity in AI outputs. The article highlights research showing that many LLMs converge on similar answers due to similar training data and methods.