associative-reasoning

Tag

Cards List
#associative-reasoning

From Reasoning Depth to Reasoning Breadth: Evaluating Multi-Point Associative Reasoning in Large Language Models

arXiv cs.CL · 4d ago Cached

This paper introduces MPAR-Bench, a bilingual benchmark for evaluating multi-point associative reasoning in large language models, along with a perturbation suite and coarse-to-fine evaluation protocol. Results show that deeper reasoning does not automatically confer robust reasoning breadth.

0 favorites 0 likes
← Back to home

Submit Feedback