video-generation

Tag

Cards List
#video-generation

Hierarchical Denoising For Multi-Step Visual Reasoning

Hugging Face Daily Papers · 2026-07-16 Cached

HDR is a unified framework integrating hierarchical latents into causal video generation for multi-step visual reasoning, achieving better reasoning consistency, lower latency, and strong data efficiency compared to baselines.

0 favorites 0 likes
#video-generation

@PrajwalTomar_: This whole video was produced by Raft's own agents, inside Raft. Humans only directed it. Give it a watch.

X AI KOLs Following · 2026-07-15 Cached

This video was entirely produced by AI agents within Raft, with only human direction, showcasing agent-driven content creation.

0 favorites 0 likes
#video-generation

KeyFrame-Compass: Towards Comprehensive Evaluation of Keyframe-Conditioned Video Generation

Hugging Face Daily Papers · 2026-07-15 Cached

KeyFrame-Compass is a benchmark and evaluation framework for keyframe-conditioned video generation, designed to assess how well models reproduce given keyframes while maintaining video quality across diverse settings.

0 favorites 0 likes
#video-generation

From Pixels to States: Rethinking Interactive World Models as Game Engines

Hugging Face Daily Papers · 2026-07-15 Cached

This paper rethinks interactive world models as game engines by examining four key dimensions—action control, state dynamics, state-observation persistence, and real-time generation—and introduces a scalable data engine for Black Myth: Wukong with over 90 hours of gameplay data to advance state-aware game world modeling.

0 favorites 0 likes
#video-generation

VideoRAE: Taming Video Foundation Models for Generative Modeling via Representation Autoencoders

Hugging Face Daily Papers · 2026-07-15 Cached

This paper introduces VideoRAE, a representation autoencoder that leverages frozen video foundation models to create compact, reconstruction-capable, and generation-friendly video latents. It achieves state-of-the-art results on UCF-101 with faster convergence than competing autoencoders.

0 favorites 0 likes
#video-generation

@omarsar0: Most world models fall apart after a few seconds. Common failure modes include texture smearing, warped geometry, and s…

X AI KOLs Timeline · 2026-07-14 Cached

LingBot-World 2.0 achieves stable 720p 60fps world model simulation for up to an hour, overcoming common failure modes like texture smearing and warped geometry.

0 favorites 0 likes
#video-generation

A new, state-of-the-art, agentic pipeline for easy Music Video creation

Reddit r/singularity · 2026-07-14

A new state-of-the-art agentic pipeline has been introduced for easy music video creation, leveraging AI to streamline the process.

0 favorites 0 likes
#video-generation

What is the model that generated this video? It's genuinely impressive

Reddit r/singularity · 2026-07-14

A user expresses admiration for a video generated by an AI model and asks which model was used, implying a significant advancement in AI video generation.

0 favorites 0 likes
#video-generation

Video Generators as General-Purpose Vision Models (8 minute read)

TLDR AI · 2026-07-14 Cached

GenCeption repurposes pre-trained video generative models into a single unified feed-forward vision model that achieves state-of-the-art performance across multiple tasks with exceptional data efficiency, marking a shift toward general-purpose visual intelligence.

0 favorites 0 likes
#video-generation

Video-generation startup PixVerse raises $439M, valuation soars past $2B

TechCrunch AI · 2026-07-14 Cached

Singapore-based video-generation startup PixVerse raised $439 million in a Series C extension, pushing its valuation past $2 billion. The company plans to expand its world model offerings and reach global customers.

0 favorites 0 likes
#video-generation

@svpino: This model can generate coherent 1+ hour videos across multiple scenes without skipping a beat. I read their paper so y…

X AI KOLs Following · 2026-07-13 Cached

This tweet explains LingBot-World-Infinity, an open-weight video generation model that uses a training technique to recover from errors, enabling coherent hour-long videos across multiple scenes.

0 favorites 0 likes
#video-generation

An open model predicting a robot's actions from a control signal. The corner panels are the action and hand pose it was given, everything else is imagined. Is this a world model, or just a video generator?

Reddit r/singularity · 2026-07-12

An open model that predicts a robot's actions from a control signal, raising questions about whether it constitutes a world model or just a video generator.

0 favorites 0 likes
#video-generation

Cseti/LTX2.3-22B_IC-LoRA-CrossView-Prompt

Hugging Face Models Trending · 2026-07-11 Cached

A proof-of-concept In-Context LoRA adapter for LTX-Video 2.3 that re-renders video scenes from new camera angles using a fixed vocabulary prompt, trained on synthetic multi-view data.

0 favorites 0 likes
#video-generation

Wan-AI/Wan-Dancer-14B

Hugging Face Models Trending · 2026-07-10 Cached

Wan-Dancer is a hierarchical framework for generating long-duration, coherent dance videos from music, with model weights and inference code released on Hugging Face.

0 favorites 0 likes
#video-generation

Video Generation Models are General-Purpose Vision Learners

Hugging Face Daily Papers · 2026-07-10 Cached

This paper proposes that large-scale text-to-video generation can serve as a powerful pre-training paradigm for computer vision, introducing GenCeption which achieves state-of-the-art performance across diverse vision tasks with high data efficiency and emergent generalization to unseen domains.

0 favorites 0 likes
#video-generation

@Saccc_c: Grok is pretty impressive! Tested another video, theme 'World Cup Quarterfinal Preview' — it's really fast, check it out.

X AI KOLs Following · 2026-07-09 Cached

The user tested Grok's video generation capability, showcasing a World Cup quarterfinal preview video, and noted significant improvement, fast generation speed, and impressive visual effects.

0 favorites 0 likes
#video-generation

Dynamic-in-Few-Step: Unifying Dynamic Computation and Few-Step Distillation for Efficient Video Generation

arXiv cs.AI · 2026-07-09 Cached

The paper proposes a post-training acceleration framework for video diffusion models that integrates dynamic structural sparsification with few-step distillation, achieving significant speedup while maintaining quality.

0 favorites 0 likes
#video-generation

OPSD-V: On-Policy Self-Distillation for Post-Training Few-Step Autoregressive Video Generators

Hugging Face Daily Papers · 2026-07-09 Cached

OPSD-V improves few-step autoregressive video diffusion models by using real long-video data as temporal context during training, providing dense trajectory-level supervision that enhances visual quality and motion dynamics without altering inference mechanisms.

0 favorites 0 likes
#video-generation

OpenCoF: Learning to Reason Through Video Generation

Hugging Face Daily Papers · 2026-07-09 Cached

OpenCoF introduces a reasoning video dataset and a fine-tuned video generation model that improves temporal reasoning through diverse supervision and explicit reasoning tokens, showing significant gains on four video reasoning benchmarks.

0 favorites 0 likes
#video-generation

robbyant/lingbot-video-moe-30b-a3b

Hugging Face Models Trending · 2026-07-08 Cached

LingBot-Video is the first open-source large-scale MoE video generation model for embodied intelligence, featuring efficient MoE architecture, massive embodied data training, and multi-reward system for high aesthetics, physical rationality, and task completion.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback