Tag
This paper identifies a confound called 'instruction leakage' in goal-conditioned compact world models for spatial relations, where the model achieves high accuracy by transcribing the instruction rather than genuinely grounding the relation. The authors propose a detection protocol and a fix that removes the goal from the dynamics and supervises the read path, recovering genuine grounding.