medical

Tag

Cards List
#medical

FOCUS: Decoupling Expert Personas in LLMs to Enhance Domain Expert Capabilities

arXiv cs.CL · 2026-08-07 Cached

Presents FOCUS, a fine-tuning method that decouples expert personas in LLMs via orthogonal decomposition and an expert gating module, improving domain-specific task accuracy across financial, legal, and medical benchmarks.

0 favorites 0 likes
#medical

Medical model: Reasoning-Medical-27B (Qwen3.6-27B finetune)

Reddit r/LocalLLaMA · 2026-07-28

Reasoning-Medical-27B is a fine-tuned Qwen3.6-27B model for advanced medical reasoning, trained on 370k Q&A examples with Chain-of-Thought reasoning using GRPO and Unsloth optimization.

0 favorites 0 likes
#medical

MedLoCoMo: A Long-Context Multi-Session Medical Dialogue Benchmark for Large Language Models

arXiv cs.AI · 2026-07-28 Cached

MedLoCoMo is a new benchmark for evaluating LLMs on long-context, multi-session medical dialogue reasoning, constructed from MIMIC-IV data. It tests single-admission, cross-admission, and adversarial unanswerable questions, revealing that cross-admission reasoning remains challenging even for models with long context windows.

0 favorites 0 likes
#medical

Couple pay >$800k for a gene-editing therapy for their daughter. She died.

Hacker News Top · 2026-07-23

A couple paid over $800,000 for a gene-editing therapy for their daughter, who subsequently died.

0 favorites 0 likes
#medical

Safety That Does Not Transfer: Cross-Lingual Clinical Correctness Drift in Deployable Medical Language Models

arXiv cs.CL · 2026-07-21 Cached

This paper investigates cross-lingual clinical correctness drift in medical language models, finding that locally deployable models show significant safety degradation when queried in Hausa compared to English, while frontier models maintain competence, highlighting a critical gap in safety evaluation for low-resource settings.

0 favorites 0 likes
#medical

Evaluating Large Language Models on Misconceptions in Multi-Turn Medical Conversations

arXiv cs.CL · 2026-07-15 Cached

The paper introduces ThReadMed-QA, a multi-turn medical dialogue dataset, and evaluates five LLMs on correcting patient misconceptions, finding substantial degradation over subsequent turns.

0 favorites 0 likes
#medical

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation

arXiv cs.AI · 2026-07-13 Cached

MedRealMM is a new multimodal benchmark for Chinese online medical consultation, built from real-world patient-doctor interactions, evaluating LLMs on next-response generation with clinical rubrics.

0 favorites 0 likes
#medical

Damn muse 1.1 now better than fable 5 for medical and legal use ??

Reddit r/singularity · 2026-07-09

Damn muse 1.1 is claimed to outperform fable 5 in medical and legal use cases.

0 favorites 0 likes
#medical

OpenMed 1.8: Apache-2.0 clinical de-identification that runs fully local, now on Android, iOS, and in the browser. 400+ open issues if you want in on 1.9

Reddit r/LocalLLaMA · 2026-07-09

OpenMed 1.8 is an Apache-2.0 toolkit for clinical de-identification that runs entirely locally, with new support for Android, iOS, and browser platforms, and invites community contributions for version 1.9.

0 favorites 0 likes
#medical

MedPMC: A Systematic Framework for Scaling High-Fidelity Medical Multimodal Data for Foundation Models

Hugging Face Daily Papers · 2026-07-08 Cached

MedPMC is an automated framework that transforms medical literature into high-fidelity multimodal data for foundation models, achieving significant improvements across multiple benchmarks and clinical settings.

0 favorites 0 likes
#medical

MedCalc-Pro: Solving Complex Medical Calculations with LLM Agents

arXiv cs.AI · 2026-07-07 Cached

The paper introduces MedCalc-Pro, a new benchmark for evaluating LLMs in complex medical calculations involving single, multi, and nested calculator settings, along with an agent framework that improves performance through multi-tool selection and structured validation.

0 favorites 0 likes
#medical

@rohanpaul_ai: https://x.com/rohanpaul_ai/status/2074005084661485771

X AI KOLs Following · 2026-07-06 Cached

A tweet shares a Nature Medicine article, but the linked content is an error page due to browser issues.

0 favorites 0 likes
#medical

FaithMed: Training LLMs For Faithful Evidence-Based Medical Reasoning

arXiv cs.CL · 2026-07-03 Cached

FaithMed is a framework that trains LLMs for faithful evidence-based medical reasoning by integrating clinician-designed rubrics with reinforcement learning using step-level process reward assignment, achieving significant improvements over baselines on multiple medical benchmarks.

0 favorites 0 likes
#medical

Cross-Domain Feature Expansion for Tabular Medical Data via Knowledge Graphs Injection

arXiv cs.AI · 2026-07-01 Cached

This paper introduces MedKGTab, a knowledge-injected framework that uses biomedical knowledge graphs to expand cross-domain features in tabular medical data, addressing data scarcity by generating high-fidelity biomedical profiles.

0 favorites 0 likes
#medical

Discrete Diffusion Language Models for Interactive Radiology Report Drafting

Hugging Face Daily Papers · 2026-07-01 Cached

This paper adapts a mixture-of-experts diffusion language model, DiffusionGemma-26B, for interactive radiology report drafting, showing it matches or exceeds autoregressive models in medical VQA with 3.5-4.4x faster decoding and bidirectional infill capabilities.

0 favorites 0 likes
#medical

IMCBench: A benchmark for multimodal LLMs in Image-grounded Medical Conversations

arXiv cs.AI · 2026-06-30 Cached

IMCBench is a new benchmark for evaluating multimodal LLMs on image-grounded medical conversations, pairing clinical images with synthetic patient profiles. Evaluations across safety, accuracy, and uncertainty show that even strong models like Claude Opus 4.6 have safety issues, highlighting the need for multi-dimensional evaluation.

0 favorites 0 likes
#medical

TriageRA-CCF: Source-Side Clinical Confidence and Coverage Signals for Adaptive Rank Budgeting in Medical LLMs

arXiv cs.CL · 2026-06-30 Cached

This paper proposes TriageRA-CCF, a method for adaptive rank budgeting in LoRA for medical question answering. It uses source-side signals (base-model confidence, clinical coverage, counterfactual proxy) to dynamically choose rank budgets, achieving modest accuracy gains on Qwen3-8B and Llama3.1-8B.

0 favorites 0 likes
#medical

@MaziyarPanahi: A year ago, OpenMed didn't exist. Today: 340M model downloads. 1,500+ open medical models, all Apache 2.0. 650+ run on …

X AI KOLs Following · 2026-06-29 Cached

A year after its inception, OpenMed has achieved 340 million model downloads, offering over 1,500 open medical models under Apache 2.0, with 650+ capable of running on-device on iPhones.

0 favorites 0 likes
#medical

Streaming medical STT running locally on a MacBook

Reddit r/LocalLLaMA · 2026-06-26

Describes a medical speech-to-text system that runs locally on a MacBook, enabling streaming transcription without cloud dependency.

0 favorites 0 likes
#medical

Explainable Ensemble-Based Machine Learning Models for Detecting the Presence of Cirrhosis in Hepatitis C Patients

arXiv cs.AI · 2026-06-26 Cached

This paper applies ensemble machine learning models (Random Forest, Gradient Boosting, XGBoost, Extra Trees) to detect cirrhosis in hepatitis C patients using 28 features from 2038 Egyptian patients. The Extra Trees model achieved 96.92% accuracy with only 16 features, outperforming other models.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback