video-action-model

Tag

Cards List
#video-action-model

Zero-WAM: In-Context World-Action Modeling from Human Videos for Open-Ended Task Generalization

Hugging Face Daily Papers · 2026-08-26 Cached

Zero-WAM is a causal video-action model that enables zero-shot robotic manipulation of unseen tasks by conditioning on in-context human video guidance, with the HumanGen dataset and a future-chunk prediction objective to improve generalization.

0 favorites 0 likes
#video-action-model

@rohanpaul_ai: Most video-action robot models are a content-creation video generator with an action module attached. LingBot-VA 2.0 fr…

X AI KOLs Timeline · 2026-07-13 Cached

LingBot-VA 2.0 is a video-action foundation model trained from scratch for robot control, achieving 225 Hz closed-loop execution with 13B parameters (1.9B active per token) and outperforming prior models on RoboTwin 2.0.

0 favorites 0 likes
#video-action-model

Any thoughts on this robot picking objects off a moving conveyor belt at 1x?

Reddit r/artificial · 2026-07-10

A robot using the LingBot-VA 2.0 video-action model picks objects off a moving conveyor belt in real-time at 1x speed, predicting future movements rather than reacting only to the current frame.

0 favorites 0 likes
← Back to home

Submit Feedback