structured-hypothesis-embeddings

Tag

Cards List
#structured-hypothesis-embeddings

Detecting an Effect Is Not Learning to Act on It: A Reward-SNR Floor for LLM Acquisition Agents

arXiv cs.LG ↗ · 2026-08-12 Cached

The paper argues that detecting an average effect of an acquired LLM-derived signal is not the same as learning per-instance acquisition policies, and establishes a reward-SNR floor (ρ* ≈ 2.8/√N) below which offline routing is impossible. It introduces Structured Hypothesis Embeddings (SHE) and shows across three datasets that learned per-example acquisition collapses below this floor, recommending design-time regime gates instead.

0 favorites 0 likes
← Back to home

Submit Feedback