prompt-embedding

Tag

Cards List
#prompt-embedding

Prompt Embedding Probes (PEP): Hallucination Detection in LLMs from Hidden States

arXiv cs.CL · 3d ago Cached

This paper introduces Prompt Embedding Probes (PEP), a parameter-efficient extension of linear probes that uses learnable prompt embeddings on hidden states to detect hallucinations in frozen LLMs. Evaluations on TriviaQA, GSM8K, and MedQA with Qwen3 models show improvements over standard linear probes, including in pre-generation and cross-model settings.

0 favorites 0 likes
← Back to home

Submit Feedback