activevision

Tag

Cards List
#activevision

GPT-5.5 Scores 10.6% on ActiveVision, Humans Hit 96.1% [R]

Reddit r/MachineLearning · 5d ago

A new arXiv paper introduces the ActiveVision benchmark designed to test repeated visual perception, finding that frontier vision models like GPT-5.5 and Claude Fable 5 score only 10.6% and 3.5% respectively, while humans achieve 96.1%.

0 favorites 0 likes
← Back to home

Submit Feedback