video-generation

Tag

Cards List
#video-generation

Sol-Attn: Accelerating Video Generation Inference via On-the-Fly Attention Sparsification

Hugging Face Daily Papers · 2026-07-27 Cached

Sol-Attn introduces a training-free method to sparsify attention for video generation inference, achieving over 2x speedup by dynamically selecting key-value blocks during online softmax with minimal quality loss.

0 favorites 0 likes
#video-generation

FilmBench: A Film-Grade Benchmark for Cinematic Video Generation

Hugging Face Daily Papers · 2026-07-27 Cached

FilmBench is a new benchmark for cinematic video generation, using prompts derived from award-winning films across 20 genres and a three-level cinematic taxonomy with 35+ sub-metrics. It includes an open-source automatic evaluation agent (FilmOps) that reproduces human model rankings with high correlation, revealing significant gaps in dynamic aesthetics and multi-shot performance compared to prior web-style benchmarks.

0 favorites 0 likes
#video-generation

Nvidia's New Long-Form Video Generation (12 minute read)

TLDR AI · 2026-07-27 Cached

NVIDIA Research introduces SANA-Video 2.0, a hybrid video diffusion transformer that generates high-quality 720p video on a single GPU, achieving up to 120× speedup over Wan 2.2-14B via hybrid linear-softmax attention and block attention residuals.

0 favorites 0 likes
#video-generation

@seekjourney: Regret finding out so late, discovered a treasure repository through my GitHub radar: video-shotcraft - 106 shot recipes, 162 styles, 161 dynamic previews, and also organized storyboards, camera movements, rhythm, sound effects, and Remotion implementations into Agent Skills. …

X AI KOLs Timeline · 2026-07-24 Cached

A GitHub repository video-shotcraft provides 106 shot recipes, 162 styles, and 161 dynamic previews organized as Agent Skills for AI video generation using Remotion and compatible with Claude Code and Codex.

0 favorites 0 likes
#video-generation

@DanKornas: Building video-generation workflows is easier when inference, model configs, and integration paths live in one place. L…

X AI KOLs Timeline · 2026-07-24 Cached

LTX-Video is an open-source Python repository by Lightricks for generating and conditioning videos locally using LTX-Video models, with support for text/image inputs, multi-condition workflows, and integration with ComfyUI and Diffusers.

0 favorites 0 likes
#video-generation

@DanKornas: Visual AI workflows get hard to manage when every model and parameter lives behind a different script. ComfyUI is a nod…

X AI KOLs Timeline · 2026-07-23 Cached

ComfyUI is an open-source, node-based AI creation engine that lets builders and visual professionals design complex generation workflows for image, video, audio, and 3D without coding, with partial re-execution and broad model support.

0 favorites 0 likes
#video-generation

Black Forest Lab's Flux 3: Omni-modality for image, video, audio & action prediction

Reddit r/singularity · 2026-07-23

Black Forest Lab's Flux 3 is a new omni-modal AI model capable of generating and predicting images, video, audio, and actions.

0 favorites 0 likes
#video-generation

BFL Introduces FLUX 3 - multi-modal model for Image, Video and Audio

Reddit r/ArtificialInteligence · 2026-07-23

BFL has introduced FLUX 3, a multi-modal AI model capable of generating images, videos, and audio.

0 favorites 0 likes
#video-generation

seedance2 depth-video to video

Reddit r/singularity · 2026-07-23

seedance2 is an AI model that converts depth-video inputs into full video outputs.

0 favorites 0 likes
#video-generation

Closing the Loop: Training-Free Revisit Consistency for Autoregressive Generative Rendering

Hugging Face Daily Papers · 2026-07-23 Cached

This paper introduces a training-free method to improve revisit consistency in autoregressive generative rendering by using temporal and spatial correspondences from the 3D engine to maintain consistent appearance when the camera revisits locations.

0 favorites 0 likes
#video-generation

SANA-Video 2.0: Hybrid Linear Attention with Attention Residuals for Efficient Video Generation

Hugging Face Daily Papers · 2026-07-23 Cached

SANA-Video 2.0 introduces a hybrid linear-softmax attention mechanism for video diffusion transformers, achieving high-quality video generation up to 720p on a single GPU with significantly reduced latency compared to full-softmax models, while maintaining competitive VBench scores.

0 favorites 0 likes
#video-generation

GraphVid: Interactive Graph-Controllable Video Generation

Hugging Face Daily Papers · 2026-07-23 Cached

GraphVid introduces a graph-conditioned image-to-video generation model that enables interactive control through structured interaction graphs, outperforming prior methods with significant reductions in FID and FVD.

0 favorites 0 likes
#video-generation

AlayaWorld is a full-stack, open-source video world model that supports 720p, 24 FPS streaming video generation with camera control

Reddit r/singularity · 2026-07-22

AlayaWorld is a full-stack, open-source video world model capable of generating 720p, 24 FPS streaming video with camera control, enabling dynamic scene creation.

0 favorites 0 likes
#video-generation

Self Gradient Forcing: Native Long Video Extrapolation

Papers with Code Trending · 2026-07-22 Cached

Proposes Self Gradient Forcing (SGF), a two-pass training strategy for autoregressive video diffusion models that provides missing supervision for writing useful context memory, enabling strong long-video extrapolation even from short training windows.

0 favorites 0 likes
#video-generation

Japan accelerates video generation with new series of anime generation models.

Reddit r/singularity · 2026-07-20 Cached

Japan's AIdea Labs released AnimeGen, a free AI model for anime-style video generation, capable of text-to-video and image-to-video, with commercial use allowed.

0 favorites 0 likes
#video-generation

ShotPlan: Cinematic Video Generation with Learnable Planning Token

Hugging Face Daily Papers · 2026-07-20 Cached

ShotPlan introduces learnable planning tokens with fractional temporal rotary position embeddings for cinematic multi-shot video generation, enabling explicit shot-level planning and achieving superior inter-shot consistency.

0 favorites 0 likes
#video-generation

HOMIE: Human-object Centric Video Personalization via Multimodal Intelligent Enchancement

Hugging Face Daily Papers · 2026-07-20 Cached

HOMIE is a framework for human-object centric video personalization that integrates MLLM features to improve subject fidelity and interaction patterns, achieving state-of-the-art performance.

0 favorites 0 likes
#video-generation

HarmoHOI: Harmonizing Appearance and 3D Motion for Multi-view Hand-Object Interaction Synthesis

Hugging Face Daily Papers · 2026-07-19 Cached

HarmoHOI is a unified diffusion framework that jointly generates synchronized multi-view hand-object interaction videos and globally aligned 3D point tracks, achieving state-of-the-art performance in visual quality, motion plausibility, and multi-view geometric consistency.

0 favorites 0 likes
#video-generation

FVAttn: Adaptive Sparse Attention with Runtime Load Balancing for Video Generation

Hugging Face Daily Papers · 2026-07-17 Cached

FVAttn is a training-free sparse attention system that uses runtime load balancing to improve distributed execution efficiency of adaptive sparse attention under multi-GPU sequence parallelism for video generation, achieving up to 4.41x attention speedup and 2.11x inference speedup over FlashAttention on step-distilled Wan2.2 I2V while maintaining competitive video quality.

0 favorites 0 likes
#video-generation

Apple-π: Benchmarking Thinking with Video Towards Law-Grounded Physical Intelligence

Hugging Face Daily Papers · 2026-07-17 Cached

Apple-π is a benchmark that evaluates video generation models on their ability to reason about physical laws through a three-stage protocol: perception, formulation, deduction. It includes 400 videos covering classical mechanics tasks and reveals current models fall short of reliable law-grounded world simulation.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback