associative-reasoning

标签

Cards List
#associative-reasoning

From Reasoning Depth to Reasoning Breadth: Evaluating Multi-Point Associative Reasoning in Large Language Models

arXiv cs.CL · 4天前 缓存

This paper introduces MPAR-Bench, a bilingual benchmark for evaluating multi-point associative reasoning in large language models, along with a perturbation suite and coarse-to-fine evaluation protocol. Results show that deeper reasoning does not automatically confer robust reasoning breadth.

0 人收藏 0 人点赞
← 返回首页

提交意见反馈