video-models

Tag

Cards List
#video-models

HappyWorld-Bench

Hugging Face Daily Papers ↗ · 5d ago Cached

HappyWorld-Bench is a comprehensive benchmark that evaluates the reliability of world models under interaction and modification across video, spatial, and embodied tracks.

0 favorites 0 likes
#video-models

@marclou: 1:05:11 I finished 1st on my age group and 5th in the male category. I failed my ambitious goal of sub-60 but I was ver…

X AI KOLs Following ↗ · 6d ago Cached

Marc Lou shares his achievement in a Hyrox race, finishing 1st in his age group, and expresses gratitude to his AI sponsors including Higgsfield AI for their support.

0 favorites 0 likes
#video-models

ActionSplice: In-Flight Action Editing for Interactive World Models

Hugging Face Daily Papers ↗ · 2026-09-08 Cached

ActionSplice introduces counterfactual state transport to splice revised actions into chunk-autoregressive video world models without replaying completed evaluations, improving fidelity and inference speed.

0 favorites 0 likes
#video-models

Principia: Relational Physics Tests for Video Models

Hugging Face Daily Papers ↗ · 2026-09-03 Cached

Principia is a benchmark that evaluates video models on Newtonian physics using relational consistency between paired objects, revealing significant gaps in current models' physical reasoning.

0 favorites 0 likes
#video-models

@CoorsLightCEO: Video models are 100% solved

X AI KOLs Timeline ↗ · 2026-08-28

A tweet from the Coors Light CEO asserts that video models in AI are fully solved, presenting a bold claim about the current state of AI development.

0 favorites 0 likes
#video-models

Visual Prompts in Video Models (8 minute read)

TLDR AI ↗ · 2026-07-30 Cached

Visual prompt engineering (VIPE) automatically modifies task images to improve video model reasoning performance, often more effective than text-based prompting or test-time scaling.

0 favorites 0 likes
#video-models

Visual prompt engineering for video models

Hugging Face Daily Papers ↗ · 2026-07-28 Cached

This paper introduces Visual Prompt Engineering (VIPE), a method that automatically modifies task images to improve video model performance, showing it can be more effective than text-based prompt engineering or test-time scaling.

0 favorites 0 likes
#video-models

@DengHokin: I am super excited to share that I launch a weekly Video Model Journal Club. Every week we pick one paper and go deep, …

X AI KOLs Timeline ↗ · 2026-06-16 Cached

The author launches a weekly Video Model Journal Club covering video generation, world models, physical reasoning, diffusion, flow matching, etc. The first in-person talk will be by Yilun Du on Embodied Reasoning with World Models.

0 favorites 0 likes
#video-models

@rohanpaul_ai: Most video models look better than they understand and Video quality is only the easiest thing to notice. LongCat just …

X AI KOLs Following ↗ · 2026-06-02 Cached

LongCat released WBench, a benchmark for video world models that tests control, memory, instruction-following, and physical plausibility across 289 cases and 20 models, finding that no model excels in all dimensions, highlighting the gap between video quality and true world simulation.

0 favorites 0 likes
#video-models

@altryne: Omni isn't Veo/Seddance! I had the awesome pleasure to have breakfast with @JeffDean during #googleio and ask him about…

X AI KOLs Following ↗ · 2026-05-21 Cached

A tweet thread reveals that Google's Omni is distinct from video models like Veo and Seedance, with DeepMind's Jeff Dean clarifying its unique input/output capabilities, described as a transformative AR filter for video.

0 favorites 0 likes
← Back to home

Submit Feedback