speed-accuracy

Tag

Cards List
#speed-accuracy

LocateAnything: Fast and High-Quality Vision-Language Grounding with Parallel Box Decoding

Hugging Face Daily Papers · 2026-05-26 Cached

LocateAnything proposes Parallel Box Decoding for unified visual grounding and object detection, decoding geometric elements as atomic units to improve throughput and localization accuracy, supported by a large-scale dataset of 138M samples.

0 favorites 0 likes
← Back to home

Submit Feedback