neural-code

Tag

Cards List
#neural-code

Polar probe linearly decodes semantic structures from LLMs

arXiv cs.CL · 2026-05-15 Cached

This paper proposes a Polar Probe that linearly recovers semantic structures from LLM activations by representing entity relations through distance and direction in a learned subspace. Testing across arithmetic, visual scenes, family trees, metro maps, and social interactions shows the code emerges in middle layers, generalizes to new entities, and causally influences model predictions.

0 favorites 0 likes
← Back to home

Submit Feedback