medical-benchmarks

Tag

Cards List
#medical-benchmarks

Structure Over Scale: Schema-Constrained Causal Graphs for RAG

arXiv cs.AI · yesterday Cached

This paper introduces HCG-RAG, which uses schema-constrained causal graphs for retrieval-augmented generation, achieving 3-20x fewer nodes and 8x-135x fewer LLM calls while matching or exceeding baseline answer quality on medical benchmarks.

0 favorites 0 likes
#medical-benchmarks

Concretized Proposition Prompting Resolves Composition-Knowledge Dichotomy in Large Language Models

arXiv cs.AI · 2026-07-10 Cached

This paper proposes Concretized Proposition Prompting (CPP), a framework that resolves the composition-knowledge dichotomy in LLMs by explicitly concretizing propositions relevant to questions, significantly enhancing reasoning performance especially in medical and math benchmarks.

0 favorites 0 likes
#medical-benchmarks

mamabench and mamaretrieval: Benchmarks for Evaluating Medical Retrieval-Augmented Generation in Maternal, Neonatal, and Reproductive Health

arXiv cs.CL · 2026-06-30 Cached

This paper introduces MamaBench and MamaRetrieval, two benchmarks for evaluating medical retrieval-augmented generation in maternal, neonatal, and reproductive health, addressing gaps in existing QA and retrieval datasets.

0 favorites 0 likes
← Back to home

Submit Feedback