medical-ai

Tag

Cards List
#medical-ai

MedBench v5: A Dynamic, Process-Oriented, and Hallucination-Aware Benchmark for Clinical Multimodal Models

arXiv cs.CL · 2026-06-24 Cached

MedBench v5 is a dynamic, process-oriented benchmark for clinical multimodal models that integrates hallucination detection and stress testing, moving beyond static QA to evaluate reasoning and stability under information-flow stressors.

0 favorites 0 likes
#medical-ai

REVEAL++: Differentiable Phenotypic Grouping for Vision-Language Retinal Modeling of Alzheimer's Disease Risk

arXiv cs.AI · 2026-06-20 Cached

This paper introduces REVEAL++, a differentiable phenotypic grouping method for vision-language contrastive learning, applied to retinal fundus images and clinical risk narratives for Alzheimer's disease risk prediction, outperforming discrete grouping baselines.

0 favorites 0 likes
#medical-ai

Using AI to help physicians diagnose rare genetic diseases affecting children

Reddit r/singularity · 2026-06-18 Cached

Researchers from Boston Children's Hospital, Harvard, and OpenAI used the OpenAI o3 Deep Research reasoning model to reanalyze 376 unsolved rare disease cases, leading to diagnoses in 18 additional cases (4.8% yield) after expert review and clinical confirmation. The study, published in NEJM AI, demonstrates how AI-assisted workflows can help experts revisit difficult cases as scientific knowledge evolves.

0 favorites 0 likes
#medical-ai

@OpenAI: Rare disease diagnosis is challenging, as sequencing can surface millions of variants, and medical knowledge changes co…

X AI KOLs · 2026-06-18 Cached

OpenAI highlights how o3 Deep Research can aid rare disease diagnosis by integrating clinical features, inheritance patterns, variant evidence, and scientific literature into actionable hypotheses for specialists.

0 favorites 0 likes
#medical-ai

@gregisenberg: For all the people that say that you can’t build an important business unless you raise venture capital Midjourney is b…

X AI KOLs Following · 2026-06-18 Cached

Midjourney announces a new division called 'Midjourney Medical', highlighting its bootstrapped success without venture capital.

0 favorites 0 likes
#medical-ai

Improving health intelligence in ChatGPT

OpenAI Blog · 2026-06-18 Cached

OpenAI announces significant improvements in health-related responses within ChatGPT using GPT-5.5 Instant, achieving accuracy comparable to frontier models and reducing factuality issues by 71% through physician-led evaluations.

0 favorites 0 likes
#medical-ai

Language Models as Interfaces, Not Oracles: A Hybrid LLM-ML System for Pediatric Appendicitis

arXiv cs.CL · 2026-06-18 Cached

This paper presents ClaMPAPP, a hybrid architecture that uses an LLM as an interface to extract features from clinical narratives, which are then passed to an XGBoost classifier for pediatric appendicitis diagnosis, demonstrating improved robustness and safety over end-to-end LLM baselines.

0 favorites 0 likes
#medical-ai

Midjourney Medical goes from generating ‘cat images’ to full-body ultrasound scans

The Verge · 2026-06-18 Cached

Midjourney CEO David Holz announced the Midjourney Scanner, a full-body ultrasound device using Butterfly Network's chips, and plans to open a spa in San Francisco for preventative scanning.

0 favorites 0 likes
#medical-ai

When, Where, and How: Adaptive Binning for Tabular Self-Supervised Learning

Hugging Face Daily Papers · 2026-06-18 Cached

This paper proposes Adaptive Binning, a learning-coupled feature-wise coarse-to-fine curriculum for tabular self-supervised learning that adaptively discretizes features, improving representations on medical datasets and establishing a unified benchmark.

0 favorites 0 likes
#medical-ai

New research shows how AMIE, our medical AI, could help manage health conditions.

Google AI Blog · 2026-06-17 Cached

Google's research shows that its medical AI, AMIE, can effectively manage health conditions over time, matching clinicians in reasoning and exceeding in plan preciseness and guideline alignment, according to a study published in Nature.

0 favorites 0 likes
#medical-ai

RubricsTree: Scalable and Evolving Open-Ended Evaluation of Personal Health Agents across Health Memory and Medical Skills

arXiv cs.CL · 2026-06-17 Cached

RubricsTree proposes a scalable, expert-aligned evaluation framework for personal health agents using over 100 atomic Boolean rubrics, achieving up to 66% relative gains on HealthBench across Gemini, GPT, and Qwen model families.

0 favorites 0 likes
#medical-ai

AIPatient Arena: EHR-grounded evaluation of large language models in end-to-end clinical consultation workflows

arXiv cs.CL · 2026-06-17 Cached

Introduces AIPatient Arena, an EHR-grounded evaluation framework for assessing LLMs across multiple dimensions of clinical competence. The study reveals strengths in interviewing and ethics but weaknesses in handling ambiguity and diagnostic accuracy.

0 favorites 0 likes
#medical-ai

Probing, Fusion, and Trustworthiness: A Systematic Evaluation of Foundation Model Representations for Multimodal Cancer Analysis

arXiv cs.LG · 2026-06-17 Cached

This paper systematically evaluates foundation model representations for multimodal cancer analysis, benchmarking unimodal and multimodal fusion strategies on real-world cohorts, and assessing trustworthiness via conformal prediction.

0 favorites 0 likes
#medical-ai

Configurable Clinical Information Extraction with Agentic RAG: What Works, What Breaks, and Why

Hugging Face Daily Papers · 2026-06-17 Cached

ACIE, an agentic RAG system for clinical information extraction, achieves 96.5% acceptance rate in nuclear-medicine physicians' judgments across 7,326 instances, addressing challenges of heterogeneous patient contexts and missing metadata.

0 favorites 0 likes
#medical-ai

Built a Paninian Retrieval-Augmented Generation (PRAG) framework for safer medical AI — seeking feedback

Reddit r/artificial · 2026-06-16

The PRAG framework combines traditional RAG with a Paninian rule engine for safer medical AI, achieving a 71% reduction in unsafe answers on MedQA. It provides auditable rule traces and is open-sourced.

0 favorites 0 likes
#medical-ai

Toward Vibe Medicine: A Self-Evolving Multi-Agent Framework for Clinical Decision Support

arXiv cs.AI · 2026-06-16 Cached

This paper presents VIBEMed, a multi-agent framework with a self-evolution mechanism and safety sandbox for robust clinical decision support, integrating specialized agents for diagnosis, treatment planning, and evolving clinical knowledge over time.

0 favorites 0 likes
#medical-ai

Semantic Reasoning in Medicine: The Role of Knowledge Graphs Across Five Key Domains

arXiv cs.LG · 2026-06-16 Cached

This survey reviews the role of knowledge graphs in medicine across five key domains—clinical decision support, disease prediction, health recommender systems, precision medicine, and medical question answering—discussing applications, challenges, and future directions.

0 favorites 0 likes
#medical-ai

MedLatentDx: Latent Multi-Agent Communication for Cross-Hospital Rare-Disease Diagnosis

arXiv cs.CL · 2026-06-15 Cached

MedLatentDx proposes a latent multi-agent communication framework for cross-hospital rare-disease diagnosis, using latent KV blocks to share diagnostic evidence without exposing clinical text, and introduces the CrossRare-Bench benchmark.

0 favorites 0 likes
#medical-ai

By 2050, we may see AI assistants in every home, personalized learning for every student, advanced medical treatments, smart cities, and even human-AI collaboration on a massive scale.

Reddit r/artificial · 2026-06-11

The article envisions a future by 2050 where AI assistants are in every home, education is personalized, medical treatments are advanced, cities are smart, and human-AI collaboration is widespread.

0 favorites 0 likes
#medical-ai

Lung-R1: A Knowledge Graph-Guided LLM for Pulmonary Diagnostic Reasoning

arXiv cs.AI · 2026-06-11 Cached

The paper introduces LungKG, the first structured pulmonary knowledge graph, and Lung-R1, a LLM trained via KG-constrained reasoning and reinforcement learning for pulmonary diagnostic reasoning from EMRs. Lung-R1-14B achieves state-of-the-art performance on EMR diagnosis.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback