Tag
A robotics researcher compares current robotics approaches to the language model landscape of 2023, arguing that representation prediction (JEPA) is the most scalable method as it can leverage action-free video data like YouTube, unlike other methods that require action-labeled data.