temporal-pooling

Tag

Cards List
#temporal-pooling

V-RAE: Rethinking Video Latent Spaces for Generation

Hugging Face Daily Papers ↗ · 2026-08-13 Cached

V-RAE proposes a video representation autoencoder that builds semantically organized latents from frozen vision representations to enhance video generation quality, convergence speed, and predictive modeling.

0 favorites 0 likes
← Back to home

Submit Feedback