pretrained-vlm

Tag

Cards List
#pretrained-vlm

SimpleMemVLA: A Simple but Effective Native-Video Memory for Vision-Language-Action Models

Hugging Face Daily Papers · 2026-09-02 Cached

SimpleMemVLA introduces a simple memory mechanism for Vision-Language-Action models by feeding intact timestamped video history into a pretrained VLM backbone, achieving state-of-the-art results on long-horizon manipulation tasks without dedicated memory modules.

0 favorites 0 likes
← Back to home

Submit Feedback