diffusion-model

Tag

Cards List
#diffusion-model

SeeSee21/Z-Anime

Hugging Face Models Trending · 2026-04-25 Cached

Z-Anime is a full fine-tune of Alibaba's Z-Image Base model, specialized for high-quality anime generation with support for natural language prompts and low VRAM usage.

0 favorites 0 likes
#diffusion-model

nvidia/Nemotron-Labs-Diffusion-14B

Hugging Face Models Trending · 2026-04-22 Cached

NVIDIA releases Nemotron-Labs-Diffusion, a family of tri-mode language models (3B, 8B, 14B) supporting AR, diffusion, and self-speculation decoding, achieving 2.7x-4x speed-ups over standard AR decoding.

0 favorites 0 likes
#diffusion-model

Bulding my own Diffusion Language Model from scratch was easier than I thought [P]

Reddit r/MachineLearning · 2026-04-21

Developer shares a minimalist 7.5M-parameter diffusion language model trained from scratch on Shakespeare, releasing the code as a learning resource.

0 favorites 0 likes
#diffusion-model

openbmb/VoxCPM2

Hugging Face Models Trending · 2026-04-03 Cached

VoxCPM2 is an open-source, tokenizer-free diffusion autoregressive Text-to-Speech model supporting 30 languages with 2B parameters, 48kHz audio output, and features including voice design from natural language descriptions, controllable voice cloning, and real-time streaming capabilities.

0 favorites 0 likes
#diffusion-model

k2-fsa/OmniVoice

Hugging Face Models Trending · 2026-03-30 Cached

OmniVoice is a massively multilingual zero-shot text-to-speech model supporting over 600 languages, built on a diffusion language model architecture with fast inference and voice cloning capabilities.

0 favorites 0 likes
#diffusion-model

Lightricks/LTX-2.3

Hugging Face Models Trending · 2026-03-04 Cached

Lightricks released LTX-2.3, an open-weight diffusion-based audio-video foundation model with improved quality and prompt adherence, available in multiple checkpoints including distilled and LoRA variants for local execution.

0 favorites 0 likes
#diffusion-model

circlestone-labs/Anima

Hugging Face Models Trending · 2026-01-29 Cached

Anima is a 2 billion parameter text-to-image model specialized for anime and illustration, released as open-source on Hugging Face through a collaboration between CircleStone Labs and Comfy Org.

0 favorites 0 likes
#diffusion-model

SANA-Video: Efficient Video Generation with Block Linear Diffusion Transformer

Papers with Code Trending · 2025-09-29 Cached

SANA-Video is a small diffusion model that efficiently generates high-resolution, long videos using linear attention and a constant-memory KV cache, achieving competitive performance at dramatically lower cost and faster speed compared to existing models.

0 favorites 0 likes
#diffusion-model

Hunyuan3D 2.0: Scaling Diffusion Models for High Resolution Textured 3D Assets Generation

Papers with Code Trending · 2025-01-21 Cached

Hunyuan3D 2.0 is a scalable flow-based diffusion transformer system for high-resolution textured 3D asset generation, outperforming state-of-the-art models and publicly released with code and weights.

0 favorites 0 likes
#diffusion-model

New BEST local AI image generator is here!

YouTube AI Channels · 2026-04-21 Cached

Ernie Image, a new open-source diffusion model, surpasses Zage in text rendering and prompt fidelity and can be run locally via ComfyUI with ~20 GB VRAM.

0 favorites 0 likes
← Previous
← Back to home

Submit Feedback