video-generation

Tag

Cards List
#video-generation

BFL Introduces FLUX 3 - multi-modal model for Image, Video and Audio

Reddit r/ArtificialInteligence · 2026-07-23

BFL has introduced FLUX 3, a multi-modal AI model capable of generating images, videos, and audio.

0 favorites 0 likes
#video-generation

seedance2 depth-video to video

Reddit r/singularity · 2026-07-23

seedance2 is an AI model that converts depth-video inputs into full video outputs.

0 favorites 0 likes
#video-generation

Closing the Loop: Training-Free Revisit Consistency for Autoregressive Generative Rendering

Hugging Face Daily Papers · 2026-07-23 Cached

This paper introduces a training-free method to improve revisit consistency in autoregressive generative rendering by using temporal and spatial correspondences from the 3D engine to maintain consistent appearance when the camera revisits locations.

0 favorites 0 likes
#video-generation

SANA-Video 2.0: Hybrid Linear Attention with Attention Residuals for Efficient Video Generation

Hugging Face Daily Papers · 2026-07-23 Cached

SANA-Video 2.0 introduces a hybrid linear-softmax attention mechanism for video diffusion transformers, achieving high-quality video generation up to 720p on a single GPU with significantly reduced latency compared to full-softmax models, while maintaining competitive VBench scores.

0 favorites 0 likes
#video-generation

GraphVid: Interactive Graph-Controllable Video Generation

Hugging Face Daily Papers · 2026-07-23 Cached

GraphVid introduces a graph-conditioned image-to-video generation model that enables interactive control through structured interaction graphs, outperforming prior methods with significant reductions in FID and FVD.

0 favorites 0 likes
#video-generation

AlayaWorld is a full-stack, open-source video world model that supports 720p, 24 FPS streaming video generation with camera control

Reddit r/singularity · 2026-07-22

AlayaWorld is a full-stack, open-source video world model capable of generating 720p, 24 FPS streaming video with camera control, enabling dynamic scene creation.

0 favorites 0 likes
#video-generation

Self Gradient Forcing: Native Long Video Extrapolation

Papers with Code Trending · 2026-07-22 Cached

Proposes Self Gradient Forcing (SGF), a two-pass training strategy for autoregressive video diffusion models that provides missing supervision for writing useful context memory, enabling strong long-video extrapolation even from short training windows.

0 favorites 0 likes
#video-generation

Japan accelerates video generation with new series of anime generation models.

Reddit r/singularity · 2026-07-20 Cached

Japan's AIdea Labs released AnimeGen, a free AI model for anime-style video generation, capable of text-to-video and image-to-video, with commercial use allowed.

0 favorites 0 likes
#video-generation

ShotPlan: Cinematic Video Generation with Learnable Planning Token

Hugging Face Daily Papers · 2026-07-20 Cached

ShotPlan introduces learnable planning tokens with fractional temporal rotary position embeddings for cinematic multi-shot video generation, enabling explicit shot-level planning and achieving superior inter-shot consistency.

0 favorites 0 likes
#video-generation

HOMIE: Human-object Centric Video Personalization via Multimodal Intelligent Enchancement

Hugging Face Daily Papers · 2026-07-20 Cached

HOMIE is a framework for human-object centric video personalization that integrates MLLM features to improve subject fidelity and interaction patterns, achieving state-of-the-art performance.

0 favorites 0 likes
#video-generation

HarmoHOI: Harmonizing Appearance and 3D Motion for Multi-view Hand-Object Interaction Synthesis

Hugging Face Daily Papers · 2026-07-19 Cached

HarmoHOI is a unified diffusion framework that jointly generates synchronized multi-view hand-object interaction videos and globally aligned 3D point tracks, achieving state-of-the-art performance in visual quality, motion plausibility, and multi-view geometric consistency.

0 favorites 0 likes
#video-generation

FVAttn: Adaptive Sparse Attention with Runtime Load Balancing for Video Generation

Hugging Face Daily Papers · 2026-07-17 Cached

FVAttn is a training-free sparse attention system that uses runtime load balancing to improve distributed execution efficiency of adaptive sparse attention under multi-GPU sequence parallelism for video generation, achieving up to 4.41x attention speedup and 2.11x inference speedup over FlashAttention on step-distilled Wan2.2 I2V while maintaining competitive video quality.

0 favorites 0 likes
#video-generation

Apple-π: Benchmarking Thinking with Video Towards Law-Grounded Physical Intelligence

Hugging Face Daily Papers · 2026-07-17 Cached

Apple-π is a benchmark that evaluates video generation models on their ability to reason about physical laws through a three-stage protocol: perception, formulation, deduction. It includes 400 videos covering classical mechanics tasks and reveals current models fall short of reliable law-grounded world simulation.

0 favorites 0 likes
#video-generation

Kimi K3 Release Video [Made with Kimi K3]

Reddit r/LocalLLaMA · 2026-07-16

Kimi K3 has been released, showcased in a video created using the model.

0 favorites 0 likes
#video-generation

Create, edit and star in videos with two Google Vids updates

Google AI Blog · 2026-07-16 Cached

Google Vids introduces Gemini Omni for generating and editing videos via natural language, and personal avatars that let users create digital versions of themselves from a selfie and voice recording without filming.

0 favorites 0 likes
#video-generation

@chetaslua: its over ........ no mcp , skills , or tool used only pure code Kimi K3 is motion graphic champion Holy Shit it coded t…

X AI KOLs Timeline · 2026-07-16 Cached

Kimi K3, a new AI model, can generate complete motion graphic videos from a single prompt without any tools or visual feedback, outperforming previous methods.

0 favorites 0 likes
#video-generation

@IndieDevHailey: Turn HTML into professional videos directly! — html-video html-video lets your Agent just write HTML + CSS + data, and render high-quality MP4 locally! It's basically the HTML version of Clipchamp! Product promos, knowledge explanations, data visualizations… all scenarios covered. Built-in 21 beautifully designed templates, AI-generated voiceover and background music, one-click export and ready to use.

X AI KOLs Timeline · 2026-07-16 Cached

html-video is an open-source tool that renders HTML+CSS+data into high-quality MP4 videos, supports AI voiceover and background music, has 21 built-in templates, and is compatible with mainstream AI agents.

0 favorites 0 likes
#video-generation

MeanFlowNFT: Bringing Forward-Process RL to Average-Velocity Generators

Hugging Face Daily Papers · 2026-07-16 Cached

MeanFlowNFT introduces a forward-process reinforcement learning method for average-velocity generators, enabling efficient alignment with human preferences while preserving fast few-step sampling. Experiments show it outperforms prior RL-tuned few-step generators on most metrics and even surpasses multi-step RL-tuned diffusion models.

0 favorites 0 likes
#video-generation

Hierarchical Denoising For Multi-Step Visual Reasoning

Hugging Face Daily Papers · 2026-07-16 Cached

HDR is a unified framework integrating hierarchical latents into causal video generation for multi-step visual reasoning, achieving better reasoning consistency, lower latency, and strong data efficiency compared to baselines.

0 favorites 0 likes
#video-generation

@PrajwalTomar_: This whole video was produced by Raft's own agents, inside Raft. Humans only directed it. Give it a watch.

X AI KOLs Following · 2026-07-15 Cached

This video was entirely produced by AI agents within Raft, with only human direction, showcasing agent-driven content creation.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback