black-box-inference

Tag

Cards List
#black-box-inference

LLMs Can See the Smoke but not the Fire: Evaluating Abductive Reasoning with Elenchos

arXiv cs.AI · 2026-07-15 Cached

The paper introduces Elenchos, a generative evaluation framework for abductive reasoning in LLMs, where models must infer hidden rule changes from behavioral differences under black-box access. It finds a detection-attribution dissociation: models detect alterations but struggle to identify the specific mutations, especially under interacting mutations.

0 favorites 0 likes
← Back to home

Submit Feedback