representation-probes

Tag

Cards List
#representation-probes

Beyond Liars' Bench: The Impact of Lie Typology, Depth, and Sparsity on Deception Detection in LLMs

arXiv cs.AI · 3d ago Cached

This paper systematically studies how lie typology, representation depth, probe expressivity, and sparse features impact deception detection in LLMs, finding that detection performance is highly dependent on training data and representation choice.

0 favorites 0 likes
← Back to home

Submit Feedback