medical

Tag

Cards List
#medical

Fast medical RAG API to give your local LLMs access to facts

Reddit r/LocalLLaMA · 2026-06-25

A free RAG API using medical Wikipedia articles is now available to provide local LLMs with accurate medical facts, as demonstrated by correcting hallucinations about Lhermitte sign.

0 favorites 0 likes
#medical

MedGuards: Multi-Agent System for Reliable Medical Error Detection and Correction

arXiv cs.CL · 2026-06-25 Cached

MedGuards proposes a multi-agent framework for detecting and correcting errors in medical text using specialized agents and confidence-guided arbitration, improving reliability without additional training. Experiments on multilingual clinical notes show significant improvements.

0 favorites 0 likes
#medical

MMed-Bench-IR: A Heterogeneous Benchmark for Multilingual Medical Information Retrieval

arXiv cs.CL · 2026-06-24 Cached

MMed-Bench-IR is a heterogeneous benchmark for multilingual medical information retrieval across six languages, evaluating cross-lingual alignment, concept discrimination, and evidence retrieval. It reveals severe performance drops for non-English queries, highlighting gaps in existing English-only evaluations.

0 favorites 0 likes
#medical

@OpenAI: Many of these cases had evaded years of expert analysis. This study suggests AI could make expert-led periodic reanalys…

X AI KOLs · 2026-06-18 Cached

This study suggests that AI can make expert-led periodic reanalysis of old medical cases more scalable, helping clinicians revisit cases as medical knowledge advances and potentially bring answers to more cases that previously evaded analysis.

0 favorites 0 likes
#medical

Towards Next-Generation Healthcare: A Survey of Medical Embodied AI for Perception, Decision-Making, and Action

arXiv cs.AI · 2026-06-16 Cached

This paper systematically surveys the core components of medical embodied AI, emphasizing the coordinated integration of perception, decision-making, and action in clinical environments, and reviews representative applications, datasets, and future research directions.

0 favorites 0 likes
#medical

MSAIC-Net: A Multi-Scale Attention and Imbalance-Aware Contrastive Network for ECG-Based Myocardial Substrate Abnormality Detection

arXiv cs.LG · 2026-06-08 Cached

Proposes MSAIC-Net, a multi-scale attention-enhanced convolutional network for detecting myocardial substrate abnormalities from ECG signals, using imbalance-aware contrastive learning and lead-wise permutation importance for interpretability.

0 favorites 0 likes
#medical

A Multi-Domain Red Teaming Framework for Safety, Robustness, and Fairness Evaluation of Medical Large Language Models

arXiv cs.CL · 2026-06-02 Cached

This paper presents a multi-domain red teaming framework for evaluating safety, robustness, and fairness of medical LLMs across 690 clinically grounded scenarios. Results show that high aggregate accuracy can mask critical failures, and hybrid evaluation with clinician oversight is necessary for credible safety assessment.

0 favorites 0 likes
#medical

Same Question, Different Source, Different Answer: Auditing Source-Dependence in Medical Multi-Source RAG

arXiv cs.CL · 2026-05-29 Cached

This paper introduces a framework for auditing source-dependence in medical multi-source RAG systems, releasing the TransplantQA benchmark, HERO-QA retrieval strategy, and a structured-output judge to measure inter-source answer relationships. It demonstrates that better retrieval reveals more disagreement than previously estimated, and argues for shifting NLP evaluation from answer correctness to inter-source relationship analysis.

0 favorites 0 likes
#medical

Augmented Equivariant Mesh Networks for Anatomical Mesh Segmentation (ICML 2026 Workshops) [R]

Reddit r/MachineLearning · 2026-05-26

Presents EAMS, a lightweight equivariant mesh segmentation framework that generalizes across anatomical tasks, showing a trade-off between equivariance and accuracy on subtle features.

0 favorites 0 likes
#medical

MedicalBench: Evaluating Large Language Models Toward Improved Medical Concept Extraction

arXiv cs.CL · 2026-05-21 Cached

MedicalBench is a new benchmark for evaluating large language models on medical concept extraction from electronic health records, focusing on implicit reasoning and evidence grounding. It includes 823 expert-annotated examples and shows that current models perform modestly, highlighting the difficulty of extracting implicitly stated medical concepts.

0 favorites 0 likes
#medical

COTCAgent: Preventive Consultation via Probabilistic Chain-of-Thought Completion

arXiv cs.CL · 2026-05-15 Cached

COTCAgent is a hierarchical reasoning framework for longitudinal electronic health records that uses a probabilistic chain-of-thought completion approach, achieving 90.47% Top-1 accuracy on a self-built dataset and outperforming existing medical agents.

0 favorites 0 likes
#medical

OpenAI for Healthcare

OpenAI Blog · 2026-01-08 Cached

OpenAI launches OpenAI for Healthcare, a suite of enterprise products including ChatGPT for Healthcare and API solutions designed to support HIPAA-compliant AI adoption across healthcare organizations. The offering features healthcare-optimized GPT-5 models, evidence-based retrieval with citations, policy integration, and workflow automation tools already deployed at major institutions like Stanford Medicine and UCSF.

0 favorites 0 likes
← Previous
← Back to home

Submit Feedback