diffusion-models

Tag

Cards List
#diffusion-models

Simplex Relaxation for Discrete Diffusion

arXiv cs.CL · 4h ago Cached

This paper introduces Simplax, an exact Dirichlet–categorical augmentation for discrete diffusion models that enriches training objectives and reverse transitions while preserving the original categorical corruption process, improving perplexity–entropy tradeoff on OpenWebText and validity on Sudoku.

0 favorites 0 likes
#diffusion-models

Generator-Guided Inverse Sampling for L\'evy-Driven Generative Models

arXiv cs.LG · 4h ago Cached

This paper studies inverse sampling for Lévy-driven generative models, proposing a structured reverse sampler that decomposes dynamics into diffusion, small jump, and large jump components, with neural networks amortizing jump rates. The method is applied to OFDM-SISO channel estimation under mixed Gaussian and impulsive noise.

0 favorites 0 likes
#diffusion-models

Hidden in Plain Sight: Diffusion-Based Unrestricted Robotic Attacks on Vision-Language-Action Models

arXiv cs.AI · 4h ago Cached

This paper proposes DURA, a diffusion-based unrestricted robotic attack that generates visually natural adversarial patches to disrupt Vision-Language-Action (VLA) models in both white-box and black-box settings, highlighting safety risks for physically deployed robotic systems.

0 favorites 0 likes
#diffusion-models

LibraSpec: Dynamic Diffusion-Based Speculative Decoding via Marginal-Gain-Driven Optimization

arXiv cs.CL · yesterday Cached

Presents LibraSpec, a training-free, plug-and-play algorithm that dynamically selects speculative decoding lengths via marginal-gain-driven optimization, achieving consistent speedups across multiple models and benchmarks.

0 favorites 0 likes
#diffusion-models

Beyond Starry Night: Shortcut-Aware Control-State Planning for Artist-Grounded Text to Image Generation

Hugging Face Daily Papers · 5d ago Cached

This paper introduces Atelier, a method that plans explicit control states before generation to prevent text-to-image models from falling back on artist-name shortcuts, and presents ArtIntentBench for evaluating artist-grounded style control.

0 favorites 0 likes
#diffusion-models

Round-Trip Consistency: Bidirectional Diffusion Models Can Predict Their Own Rollout Errors [R]

Reddit r/MachineLearning · 5d ago

This paper introduces a bidirectional latent diffusion model that steps dynamical systems forward or backward in time, using round-trip consistency as a self-supervised test-time error signal to predict rollout errors without ground truth or ensembles.

0 favorites 0 likes
#diffusion-models

Transferable Dual-Stream Representations for Mesoscale-Preserving Sea Surface Temperature Downscaling

arXiv cs.LG · 6d ago Cached

This paper introduces EddyFlow, a deep learning framework for kilometer-scale sea surface temperature downscaling that balances predictive accuracy, scale-dependent structure, and regional generalization. It achieves strong zero-shot performance and near-ideal spectral fidelity across multiple ocean regions.

0 favorites 0 likes
#diffusion-models

Scaling Inherently Interpretable Language Models

Hugging Face Daily Papers · 6d ago Cached

This paper introduces Steerling-8B, a diffusion language model trained with interpretability as a constraint, showing that interpretability improves with scale and enabling concept steering without retraining.

0 favorites 0 likes
#diffusion-models

Long-term Traffic Scene Prediction via Polynomial Representations in Autonomous Driving

arXiv cs.AI · 2026-08-05 Cached

This thesis introduces polynomial representations for long-term traffic scene prediction in autonomous driving, showing improved computational efficiency, generalization, and prediction plausibility over sequence-based baselines, validated on Argoverse 2 and Waymo Open datasets.

0 favorites 0 likes
#diffusion-models

UniWorld-View: Large-Baseline View Synthesis via Video Diffusion Models

Hugging Face Daily Papers · 2026-08-05 Cached

Introduces UniWorld-View, a unified framework for large-baseline novel view synthesis from monocular inputs, integrating occlusion-aware point cloud rendering with video diffusion models for precise camera control and geometric consistency.

0 favorites 0 likes
#diffusion-models

FairDiffuseVQVAE: Sampling-Time Fairness in Tabular Diffusion via Conditional Refinement of Vector-Quantized Latents

arXiv cs.LG · 2026-08-03 Cached

Introduces FairDiffuseVQVAE, a two-stage tabular diffusion model that achieves fairness at sampling time by conditioning on protected attributes, outperforming prior fair tabular generators on demographic parity and equalized odds.

0 favorites 0 likes
#diffusion-models

@BioSpace9: De Novo Design of Protein Switches with Diffusion-Based Ensemble Sampling

X AI KOLs Timeline · 2026-08-03 Cached

This bioRxiv preprint introduces Diff-Switch, a framework that uses diffusion-based ensemble sampling to generate conformational states for de novo protein switch design, improving the success rate of finding switch-compatible sequences.

0 favorites 0 likes
#diffusion-models

A Frozen Pixel-Space Diffusion Model Can Guide Itself with Its Own Samples

Hugging Face Daily Papers · 2026-07-31 Cached

This paper introduces Synthetic Self-Guidance (SSG), a method that attaches a lightweight prediction head to a frozen pretrained pixel-space diffusion model, using the discrepancy between intermediate and final predictions as self-guidance during sampling. It shows that model-generated samples suffice for training the head, improving FID by over 50% on several variants without classifier-free guidance and enhancing strong baselines with CFG.

0 favorites 0 likes
#diffusion-models

Scaling Properties of Text Conditioning in Visual Generation

Hugging Face Daily Papers · 2026-07-31 Cached

This paper studies empirical scaling properties for text conditioning in visual generation, showing that converged diffusion loss scales with structured language in prompts, and introduces methods to improve diffusability and promptability.

0 favorites 0 likes
#diffusion-models

Flow Map Learning via Nongradient Vector Flow

arXiv cs.LG · 2026-07-30 Cached

This paper introduces SGFlow, a method for learning flow maps for diffusion models that avoids invertibility constraints and backpropagation through model iterations, achieving competitive FID scores on CIFAR with a proven stationary-point guarantee.

0 favorites 0 likes
#diffusion-models

Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers

Hugging Face Daily Papers · 2026-07-30 Cached

This paper introduces Chimera, a hybrid visual diffusion backbone with a principled scaling recipe, combining Kimi Delta Attention, Multi-head Latent Attention, and sparse Mixture-of-Experts to efficiently handle long-context image and video generation. It also presents HeteroP, a module-wise hyperparameter transfer scheme, and Chinchilla-style scaling laws to train an 11B-parameter model with 2B activated parameters.

0 favorites 0 likes
#diffusion-models

Parallel Decoding for Video Generation (10 minute read)

TLDR AI · 2026-07-30 Cached

NVIDIA introduces Parallel Decoding Distillation (PDD) for accelerating image and video generation, enabling high-quality outputs with fewer neural function evaluations on models like LTX-2.3 and Wan2.1-14B.

0 favorites 0 likes
#diffusion-models

Neuromorphic Diffusion Language Models: Addressing Compute and Memory Bottlenecks via Sparsity and Block Denoising

arXiv cs.CL · 2026-07-29 Cached

Proposes neuromorphic masked diffusion language models (N-MDLMs) that integrate block diffusion with spike-based neuromorphic computation to improve throughput and energy efficiency by leveraging sparsity and generating multiple tokens per parameter access, analyzed via a roofline-inspired model.

0 favorites 0 likes
#diffusion-models

Steering topology distributions for unified generative design of architected metamaterials

arXiv cs.AI · 2026-07-29 Cached

Introduces GenTO, a diffusion model-based framework that steers topology distributions for unified design of architected metamaterials, achieving diverse design tasks with reusable topology priors and experimental validation.

0 favorites 0 likes
#diffusion-models

PRESTO: Prefix-Aligned Tree Drafting for Diffusion Speculative Decoding

arXiv cs.AI · 2026-07-28 Cached

PRESTO introduces a prefix-aligned tree drafting framework for diffusion speculative decoding, achieving up to 1.5x speedup on dedicated diffusion drafters and 1.12x on self-speculative diffusion LLMs.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback