medical-imaging

Tag

Cards List
#medical-imaging

Surg-UniWorld: A Unified Surgical World Model with Multimodal Control Experts

arXiv cs.AI · 10h ago Cached

Surg-UniWorld is a unified surgical world model with multimodal control experts, enabling controllable generation of coherent instrument-tissue interaction videos using edge, depth, and optical-flow inputs. It introduces a new benchmark (Cholec80-SurgWAM) and outperforms existing controllable video generation methods.

0 favorites 0 likes
#medical-imaging

@JakobWasserthal: TotalSegmentator now has an MCP server. Connect it to e.g. Codex and ask questions like: “Does this CT show signs of he…

X AI KOLs Following · 3d ago Cached

TotalSegmentator now has an MCP server, enabling AI agents like Codex to run it and answer clinical questions about CT scans, e.g., detecting hepatosplenomegaly or NAFLD/NASH.

0 favorites 0 likes
#medical-imaging

LiNC: Lightweight Noise Correction via Per-Sample Trust and Gaussian Mixture Modeling

arXiv cs.LG · 4d ago Cached

Introduces LiNC, a lightweight noise correction method that learns per-sample trust parameters to distinguish clean and noisy labels using a Gaussian Mixture Model, achieving robust accuracy gains on medical imaging datasets under high label noise.

0 favorites 0 likes
#medical-imaging

Towards Interpretable Foundation Models for Retinal Fundus Images

Hugging Face Daily Papers · 6d ago Cached

This paper proposes DualIFM, an interpretable-by-design foundation model for retinal fundus images, achieving performance comparable to RETFound with far fewer parameters while providing interpretable predictions.

0 favorites 0 likes
#medical-imaging

Loud or Silent? A Reusable Framework for Per-Modality Failure Analysis in Multimodal Clinical AI

Hugging Face Daily Papers · 2026-08-02 Cached

This paper presents a model-agnostic framework for per-modality failure analysis in multimodal clinical AI, distinguishing loud vs silent failures when a modality is dropped. Validated on planted ground truth and applied to EchoJEPA and HuBERT-ECG embeddings for LVEF prediction, it shows that dropping echo nearly doubles error.

0 favorites 0 likes
#medical-imaging

Forensic Reproducibility Audit of a Radiology Vision-Language Model Benchmark: From Intended Protocol to Released Artifact

arXiv cs.AI · 2026-07-31 Cached

This paper performs a forensic reproducibility audit of a radiology vision-language model benchmark, finding divergences between the intended protocol and released artifacts that invalidate the original claims. The authors propose a benchmark contract to expose such failure classes.

0 favorites 0 likes
#medical-imaging

@marionlepert: Catching skin cancer early is a home robotics problem. Melanoma is highly treatable when detected early, yet today’s sc…

X AI KOLs Following · 2026-07-29 Cached

Marion Lepert introduces OpenDerm, an open-source 4-DOF robot for high-resolution skin imaging to enable early melanoma detection at home, arguing that general-purpose home robots could make routine skin screening accessible.

0 favorites 0 likes
#medical-imaging

QFedPolyp: A Communication- and Inference-Efficient Federated Learning Framework for Polyp Segmentation

arXiv cs.LG · 2026-07-28 Cached

QFedPolyp proposes a federated learning framework for polyp segmentation that uses quantization-aware training to reduce communication costs and achieve faster inference while preserving privacy.

0 favorites 0 likes
#medical-imaging

Which area of healthcare will benefit most from AI in the next few years?

Reddit r/AI_Agents · 2026-07-27

Discussion on which area of healthcare will see the biggest transformation from AI in the next few years, including diagnosis, drug discovery, patient monitoring, and medical imaging.

0 favorites 0 likes
#medical-imaging

Diffusion Models in Medical Image Inpainting: Challenges, Solution Taxonomy, and Future Directions

arXiv cs.CL · 2026-07-27 Cached

A systematic review of diffusion-based methods for medical image inpainting, covering architectures, applications, datasets, and evaluation strategies, with a proposed taxonomy and identification of challenges such as lack of standardized benchmarks.

0 favorites 0 likes
#medical-imaging

GLI-AL: A Multi-Modal Glioma MRI Label Resource with Unified Anatomy-Lesion Labels

Hugging Face Daily Papers · 2026-07-27 Cached

GLI-AL is a new label resource for glioma MRI that unifies anatomy and lesion labels, expanding supervision to include healthy tissues and previously unlabeled abnormalities. It provides 1,251 label sets aligned with BraTS-GLI cases.

0 favorites 0 likes
#medical-imaging

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding

Hugging Face Daily Papers · 2026-07-27 Cached

ClinFusion is a vision-centric multimodal large language model for holistic medical understanding that unifies 2D and 3D medical image analysis using a cascaded vision encoder. It achieves state-of-the-art results on 20 out of 24 benchmarks and outperforms proprietary models like GPT-5.2 and Gemini-3-Flash on 13 out of 16 benchmarks.

0 favorites 0 likes
#medical-imaging

@freeCodeCamp: Training a medical imaging model should start with understanding the data, not just choosing an architecture. In this t…

X AI KOLs Timeline · 2026-07-24 Cached

A tutorial from freeCodeCamp teaching how to build a tumor segmentation pipeline for breast ultrasound images using MONAI, emphasizing data profiling before model selection.

0 favorites 0 likes
#medical-imaging

PathReportEval: A Systematic Benchmark for Pathology Report Generation

arXiv cs.CL · 2026-07-22 Cached

This paper introduces PathReportEval, a standardized benchmark and evaluation framework for pathology report generation from whole-slide images, including a new clinically grounded metric called Clinical Report Quality Score (CRQS) that better captures factual correctness than conventional lexical metrics.

0 favorites 0 likes
#medical-imaging

FedCC: A Low-Resource Federated Adaptation of Foundation Models for Robust Corpus Callosum localization in Fetal Ultrasound Images

arXiv cs.LG · 2026-07-22 Cached

FedCC proposes a federated learning framework combining a frozen DINOv2 backbone with a lightweight YOLO detection head and LoRA modules for robust corpus callosum localization in fetal ultrasound images, achieving strong performance with greatly reduced communication cost in a multi-center setting.

0 favorites 0 likes
#medical-imaging

qZACH-ViT: Quantization-Aware Intrinsic Explanations with Recursive Attribution-Stabilized Optimization

arXiv cs.LG · 2026-07-20 Cached

Introduces qZACH-ViT, a quantization-aware extension of ZACH-ViT with recursive intrinsic explanations, and Recursive Attribution-Stabilized Optimization (RASO) for stable attribution gradients. Achieves high prediction agreement and speedups on MedMNIST datasets after INT8 conversion.

0 favorites 0 likes
#medical-imaging

PnP-CoSMo: A Multi-Contrast MRI Reconstruction Framework based on Content/Style Modeling [R]

Reddit r/MachineLearning · 2026-07-16

PnP-CoSMo is a plug-and-play framework for multi-contrast MRI reconstruction that learns content/style models from image data, enabling reconstruction without raw k-space training data. It is generalizable across contrasts and forward operators.

0 favorites 0 likes
#medical-imaging

Beyond Backbone Backpropagation: A Decoupled Strategy for Efficient Transfer Learning

arXiv cs.LG · 2026-07-16 Cached

Proposes a decoupled training strategy that adapts normalization layers and uses precomputed features to reduce overhead in transfer learning, achieving competitive accuracy with significantly reduced training time and energy consumption.

0 favorites 0 likes
#medical-imaging

Detecting pneumonia with quantum AI

Reddit r/singularity · 2026-07-14 Cached

LMU researchers developed a Quantum Boltzmann Machine model for pneumonia detection from chest X-rays. Using fewer than 9,000 parameters (vs. millions in classical CNNs), the model achieves 84–86% accuracy, demonstrating potential advantages of quantum machine learning for small medical datasets.

0 favorites 0 likes
#medical-imaging

Indian scientists produce most detailed 3D atlas of the human brainstem

Hacker News Top · 2026-07-14 Cached

Indian scientists at IIT-Madras have created the most detailed 3D atlas of the human brainstem at cellular resolution, called Anchor, integrating over 500 tissue sections to map more than 200 cell clusters and pathways, bridging whole-brain imaging and cellular pathology.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback