clinical-reasoning

Tag

Cards List
#clinical-reasoning

Cura 1T: Specialized Model for Agentic Healthcare

arXiv cs.AI · 2d ago Cached

Cura 1T is a healthcare-specialized LLM trained via a human-gated self-evolution loop that iteratively improves on patient consultation, clinical reasoning, and agentic healthcare tasks, achieving top performance on medical benchmarks while maintaining general reasoning ability.

0 favorites 0 likes
#clinical-reasoning

Information-seeking failures of large language models in agentic clinical reasoning

arXiv cs.AI · 2026-07-14 Cached

The paper develops an agentic evaluation framework for clinical reasoning in hematologic oncology, finding that LLMs primarily fail due to systematic information-seeking deficits rather than insufficient knowledge, with error patterns resembling cognitive biases in novice clinicians.

0 favorites 0 likes
#clinical-reasoning

Evaluating Retrieval-Augmented Generation vs. Long-Context Input for Clinical Reasoning over EHRs

arXiv cs.CL · 2026-07-13 Cached

This paper evaluates retrieval-augmented generation (RAG) versus long-context prompting for clinical reasoning tasks over electronic health records, finding RAG to be token-efficient and competitive, especially for imaging extraction and antibiotic timeline reconstruction.

0 favorites 0 likes
#clinical-reasoning

Towards Precision Therapy in Hepatocellular Carcinoma: A Clinical-Reasoning LLM for Risk Stratification and Treatment Guidance

arXiv cs.AI · 2026-07-10 Cached

This paper presents HCC-STAR, a clinically aligned large language model for risk stratification and treatment guidance in hepatocellular carcinoma, aiming to improve precision therapy by leveraging electronic medical records.

0 favorites 0 likes
#clinical-reasoning

A safety-oriented hypothetico-deductive framework for AI-assisted differential diagnosis

arXiv cs.AI · 2026-07-10 Cached

AegisDx is a safety-oriented framework that uses specialized LLM components and verification gates for hypothetico-deductive clinical reasoning, improving differential diagnosis accuracy by 7-17 percentage points over standalone LLMs on medical case reports.

0 favorites 0 likes
#clinical-reasoning

CLExEval: A Human-in-the-Loop Framework for Qualitative Evaluation of LLM Clinical Reasoning

arXiv cs.CL · 2026-07-01 Cached

CLExEval introduces a human-in-the-loop framework for evaluating LLM clinical reasoning under progressive information masking, revealing failure patterns such as verbosity bias, hidden knowledge paradox, and reasoning-to-output mismatch in models like GPT-4o-mini and HuatuoGPT-o1.

0 favorites 0 likes
#clinical-reasoning

Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning

Hugging Face Daily Papers · 2026-06-30 Cached

Introduces MRPO, a reinforcement learning algorithm that uses step-wise process rewards to mitigate cascading errors in clinical multimodal reasoning, outperforming existing methods on medical VQA benchmarks.

0 favorites 0 likes
#clinical-reasoning

Experience Makes Skillful: Enabling Generalizable Medical Agent Reasoning via Self-Evolving Skill Memory

Hugging Face Daily Papers · 2026-06-08 Cached

This paper introduces SkeMex, a self-evolving framework that enhances medical agents by distilling interaction trajectories into structured skill memory, enabling better long-term clinical reasoning through context-dependent utility estimation and governance.

0 favorites 0 likes
#clinical-reasoning

ChatHealthAI: Aligning Electronic Health Record Representations with Large Language Models for Grounded Clinical Reasoning

arXiv cs.AI · 2026-06-03 Cached

ChatHealthAI is a multimodal reasoning framework that aligns structured EHR representations with a frozen LLM to enable grounded clinical reasoning while maintaining predictive performance.

0 favorites 0 likes
#clinical-reasoning

MedGuideX: Internalizing Decision Logic from Executable Guidelines into Large Language Models for Clinical Reasoning

arXiv cs.AI · 2026-05-27 Cached

MedGuideX transforms clinical practice guidelines into executable decision logic to generate factual and counterfactual QA data for training medical LLMs, achieving a 10.28% relative improvement in average accuracy across clinical reasoning benchmarks.

0 favorites 0 likes
#clinical-reasoning

SEMA-RAG: A Self-Evolving Multi-Agent Retrieval-Augmented Generation Framework for Medical Reasoning

arXiv cs.CL · 2026-05-19 Cached

SEMA-RAG is a self-evolving multi-agent RAG framework for medical question answering that decouples interpretation, exploration, and adjudication into three specialist agents, achieving significant accuracy improvements over baselines across multiple benchmarks.

0 favorites 0 likes
#clinical-reasoning

ClinSeekAgent: Automating Multimodal Evidence Seeking for Agentic Clinical Reasoning

Hugging Face Daily Papers · 2026-05-19 Cached

ClinSeekAgent is an automated agentic framework that enables large language models to actively acquire and synthesize multimodal clinical evidence from raw data sources, improving decision-making accuracy in both text-only and multimodal tasks. It introduces the ClinSeek-Bench benchmark and a distilled model ClinSeek-35B-A3B that achieves strong performance on agentic clinical reasoning.

0 favorites 0 likes
#clinical-reasoning

Checkup2Action: A Multimodal Clinical Check-up Report Dataset for Patient-Oriented Action Card Generation

arXiv cs.CL · 2026-05-13 Cached

This paper introduces Checkup2Action, a multimodal dataset and benchmark for generating patient-oriented action cards from clinical check-up reports, addressing the interpretability gap for laypersons.

0 favorites 0 likes
← Back to home

Submit Feedback