diffusion-model

Tag

Cards List
#diffusion-model

@Celeris_ai: Introducing Celeris-1 Magnus. A model built for agentic work. On τ³-bench banking, Magnus delivers 41.2% at a 55-second…

X AI KOLs Timeline · yesterday Cached

Celeris-1 Magnus is a hybrid diffusion model optimized for agentic work, achieving a 41.2% solve rate on the τ³-bench banking benchmark at a 55-second median time, outperforming models like GPT-5.6-sol.

0 favorites 0 likes
#diffusion-model

LayerRecall: A State-Conditioned Memory Router for Long-Horizon Consistency in Video Generation

Hugging Face Daily Papers · 6d ago Cached

LayerRecall improves long-video consistency in diffusion models by selectively routing historical memory to specific layers, supervised by cross-horizon prediction matching.

0 favorites 0 likes
#diffusion-model

I trained a game music generator

Reddit r/LocalLLaMA · 2026-08-23

I trained a 1.2B DiT model for instrumental game music generation, using Stable Audio's VAE and aiming to cover diverse styles. The project is open-source with a WebUI and samples available on HuggingFace.

0 favorites 0 likes
#diffusion-model

ADAPT: Physics-Aware Diffusion-based World Models for Adaptive Predictive Transferable HVAC Control

arXiv cs.AI · 2026-08-21 Cached

ADAPT is a physics-aware conditional diffusion model for HVAC control that reduces energy consumption and occupant discomfort, with robust transferability to unseen environments.

0 favorites 0 likes
#diffusion-model

Block3D: Efficient Text-to-3D Generation via Block-Wise Diffusion

Hugging Face Daily Papers · 2026-08-20 Cached

Block3D accelerates text-to-3D generation by using block-wise diffusion with confidence-guided correction to reduce inference time while preserving geometric fidelity, achieving a 5.15x speedup.

0 favorites 0 likes
#diffusion-model

Trained an diffusion model that runs on 264KB of RAM [P]

Reddit r/MachineLearning · 2026-08-18

An individual trained a diffusion model to generate 32x32 pixel images on a Shrike lite microcontroller with only 264KB RAM, experimenting with FPGA acceleration that hit memory bottlenecks, resulting in noisy but sometimes interesting outputs.

0 favorites 0 likes
#diffusion-model

incoai/Qwen3.8-27B-DFlash2

Hugging Face Models Trending · 2026-08-18 Cached

DFlash 2 is a block-diffusion draft model for speculative decoding that improves inference speed for the Qwen3.8-27B language model, offering higher acceptance length and throughput in benchmarks.

0 favorites 0 likes
#diffusion-model

@no_stp_on_snek: this week just keeps getting better.

X AI KOLs Following · 2026-08-14 Cached

Molei Tao introduces FLARE, a diffusion language model that achieves near GPT5 performance with significantly faster inference speed.

0 favorites 0 likes
#diffusion-model

UniSwap: Streaming Audio-Visual Identity Swapping for Talking Videos

Hugging Face Daily Papers · 2026-08-13 Cached

UniSwap is a new framework for joint audio-visual identity swapping in talking videos, using a unified streaming audio-visual diffusion transformer to replace appearance and vocal timbre while preserving source content and dynamics.

0 favorites 0 likes
#diffusion-model

Surg-UniWorld: A Unified Surgical World Model with Multimodal Control Experts

arXiv cs.AI · 2026-08-10 Cached

Surg-UniWorld is a unified surgical world model with multimodal control experts, enabling controllable generation of coherent instrument-tissue interaction videos using edge, depth, and optical-flow inputs. It introduces a new benchmark (Cholec80-SurgWAM) and outperforms existing controllable video generation methods.

0 favorites 0 likes
#diffusion-model

Comfy-Org/MiniMax-Music-3

Hugging Face Models Trending · 2026-08-08 Cached

This article provides instructions for placing repackaged MiniMax-Music-3 model files into ComfyUI directories for music generation.

0 favorites 0 likes
#diffusion-model

Real-time probabilistic tsunami forecasting via generative AI

arXiv cs.LG · 2026-08-06 Cached

This paper introduces a probabilistic ensemble model based on a conditional diffusion model for real-time tsunami inundation forecasting, offering uncertainty quantification in contrast to deterministic warnings. Validated with 2011 Tohoku-oki data, it demonstrates that generative AI can shift tsunami forecasting from deterministic to probabilistic approaches.

0 favorites 0 likes
#diffusion-model

Scenema Audio Comes to ComfyUI, Runs on 8GB VRAM

Reddit r/LocalLLaMA · 2026-08-05

Scenema Audio, an expressive text-to-speech model with zero-shot voice cloning, is now available as a native ComfyUI custom node, quantized to run on 8GB VRAM. The release adds inline stage direction cues, 12 preset voices, and simplifies the prompt format for ComfyUI.

0 favorites 0 likes
#diffusion-model

SynEnergy: Anomaly Semantic-Guided Diffusion for Synthetic Energy Data Generation

arXiv cs.LG · 2026-08-05 Cached

This paper introduces SynEnergy, a two-stage diffusion-based framework for generating synthetic energy consumption data while preserving rare anomalous events using heterogeneous graph-based anomaly semantic learning and anomaly semantic-guided diffusion.

0 favorites 0 likes
#diffusion-model

UniNav: A Unified World-Action Diffusion Model for Visual Navigation

arXiv cs.AI · 2026-08-05 Cached

UniNav is a unified world-action diffusion model for image-goal visual navigation that jointly predicts future visual observations and waypoint trajectories in a single diffusion process, achieving strong benchmark results with efficient inference.

0 favorites 0 likes
#diffusion-model

MBDiff: Multi-view Behavior-aware Diffusion Model for Probabilistic Utility Data Imputation

arXiv cs.LG · 2026-08-03 Cached

Presents MBDiff, a multi-view behavior-aware diffusion model for probabilistic utility data imputation that learns user behavior from global, local, and instance-level views and uses a conditional attentional denoising network. Evaluated on real utility data from Florida, it outperforms state-of-the-art baselines.

0 favorites 0 likes
#diffusion-model

Existence-Field Diffusion Model for Spatial Point Processes with Variable Cardinality

arXiv cs.LG · 2026-07-30 Cached

Proposes the existence-field diffusion model (EFDM) that jointly models spatial locations and cardinality of point sets via a unified diffusion process with existence variables, eliminating the need for explicit discrete transitions.

0 favorites 0 likes
#diffusion-model

Diffusion-Guided Search via Exponential Tilting (DiffTilt): An Application to Falsification of Safety-Critical Systems

arXiv cs.LG · 2026-07-28 Cached

This paper introduces DiffTilt, a distributional framework that exponentially tilts a diffusion model-induced joint distribution over environments and executions to efficiently discover rare safety-critical failures in autonomous and cyber-physical systems, outperforming conditional sampling strategies on ARCH-COMP benchmarks and a new tractor-trailer benchmark.

0 favorites 0 likes
#diffusion-model

Explicit Layer Modeling for Video Object Insertion and Layer Decomposition

Hugging Face Daily Papers · 2026-07-28 Cached

This paper introduces TriLayer, a large-scale video dataset with foreground-background-composite triplets, and DBL-Diffusion, a dual-branch diffusion framework for explicit layered video representation, enabling high-fidelity object insertion and layer decomposition.

0 favorites 0 likes
#diffusion-model

the diffusion versus autoregressive debate finally has a clean data point, and it points to a much narrower claim than the hype

Reddit r/artificial · 2026-07-24

The lab behind LLaDA2.2 released a diffusion model benchmarked against its own autoregressive model, showing diffusion lags on general knowledge and coding but wins on speed and agent tasks, providing a clean tradeoff data point.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback