video-qa

Tag

Cards List
#video-qa

Evidence-Backed Video Question Answering

Hugging Face Daily Papers · 2026-07-13 Cached

This paper introduces Evidence-Backed Video Question Answering (E-VQA), a new task requiring models to output both semantic answers and precise spatio-temporal evidence like tracked object segmentation masklets. The authors create a human-verified benchmark and a scalable training dataset, showing significant improvements over baselines.

0 favorites 0 likes
#video-qa

VisualClaw: A Real-Time, Personalized Agent for the Physical World

Hugging Face Daily Papers · 2026-06-15 Cached

VisualClaw is a self-evolving multimodal agent that reduces deployment costs through hybrid encoding and skill evolution, while improving video-QA accuracy across multiple benchmarks.

0 favorites 0 likes
← Back to home

Submit Feedback