video-foundation-model

Tag

Cards List
#video-foundation-model

VideoRAE: Taming Video Foundation Models for Generative Modeling via Representation Autoencoders

Hugging Face Daily Papers · 2026-07-15 Cached

This paper introduces VideoRAE, a representation autoencoder that leverages frozen video foundation models to create compact, reconstruction-capable, and generation-friendly video latents. It achieves state-of-the-art results on UCF-101 with faster convergence than competing autoencoders.

0 favorites 0 likes
#video-foundation-model

@_akhaliq: LingBot-Video is out on Hugging Face MoE-based video foundation model built for embodied intelligence 30B params, only …

X AI KOLs Following · 2026-07-08 Cached

LingBot-Video, a 30B parameter MoE-based video foundation model for embodied intelligence, has been released on Hugging Face with only 3B active parameters at inference, augmented with 70K hours of embodied data.

0 favorites 0 likes
← Back to home

Submit Feedback