Saint-lsy/Awesome-Volumetric-Radiology-AI

Volumetric Radiology AI in the Era of Multimodal Large Language Models

23

39 commits

updated Aug 24, 2026

See the code

README

Volumetric Radiology AI in the Era of Multimodal Large Language Models

arXiv:2608.20549 Download PDF

The official resource repository for our review of volumetric radiology foundation models, multimodal large language models, agentic systems, and evaluation benchmarks.

Timeline showing the co-development of volumetric radiology foundation models and agentic systems

Co-development of volumetric radiology foundation models and agentic systems.

Contents


Foundation Models

Three-stage landscape of volumetric radiology foundation models

Foundation-model landscape: volumetric self-supervised pre-training, vision-language alignment, and MLLM-based interpretation.

Self-Supervised Pre-training

MethodTitleVenueDatePaperProject
Models GenesisMedical Image Analysis21.02PaperProject
UniMiSSUniMiSS: Universal Medical Self-Supervised Learning via Breaking Dimensionality BarrierECCV2022&TPAMI21.12PaperProject
Swin-UNETRSelf-Supervised Pre-Training of Swin Transformers for 3D Medical Image AnalysisCVPR 2022CVPR 2022PaperProject
PCRLv2A Unified Visual Information Preservation Framework for Self-supervised Pre-training in Medical Image AnalysisIEEE TPAMI2023.01PaperProject
MIM-Med3DMasked Image Modeling Advances 3D Medical Image AnalysisWACV 20232023.01PaperProject
GVSLGeometric Visual Similarity Learning in 3D Medical Image Self-supervised Pre-trainingCVPR 20232023.03PaperProject
M3AEM3AE: Multimodal Representation Learning for Brain Tumor Segmentation with Missing ModalitiesAAAI 20232023.03PaperProject
HybridMIMHybridMIM: A Hybrid Masked Image Modeling Framework for 3D Medical Image SegmentationIEEE JBHI2024.04PaperProject
VoCoLarge-Scale 3D Medical Image Pre-Training With Geometric Context PriorsCVPR 2024 / IEEE TPAMI2024.06PaperProject
CDSSL-P3DCross-Dimensional Medical Self-Supervised Representation Learning Based on a Pseudo-3D TransformationMICCAI 242024.10Paper-
MDMMasked Deformation Modeling for Volumetric Brain MRI Self-Supervised Pre-TrainingIEEE TMI2024.12PaperProject
DAEDisruptive Autoencoders: Leveraging Low-level features for 3D Medical Image Pre-trainingMIDL 20242024.06PaperProject
Unified 3D MRI Representations via Sequence-Invariant Contrastive LearningarXiv2025.01PaperProject
GzPTImproving Self-Supervised Medical Image Pre-Training by Early Alignment With Human Eye Gaze InformationAAAI24/IEEE TMI2025.01PaperProject
FM-HCT3D Foundation AI Model for Generalizable Disease Detection in Head Computed TomographyarXiv2025.02PaperProject
MedVAE: Efficient Automated Interpretation of Medical Images with Large-Scale Generalizable AutoencodersMIDL 20252025.02PaperProject
MiMMiM: Mask in Mask Self-Supervised Pre-Training for 3D Medical Image AnalysisIEEE TMI2025.04Paper-
BrainMVPBrainMVP: Multi-modal Vision Pre-training for Brain MRI AnalysisCVPR 2025 (Highlight)2025.06PaperProject
SPECTREScaling Self-Supervised and Cross-Modal Pretraining for Volumetric CT TransformersCVPR 20262025.11PaperProject
SubForeHU-based Foreground Masking for 3D Medical Masked Image ModelingMICCAI252025.10PaperProject
3DINOA generalizable 3D framework and model for self-supervised learning in medical imagingnpj Digital Medicine2025.11PaperProject
TotalFMTotalFM: An Organ-Separated Framework for 3D-CT Vision Foundation Modelsarxiv2026.01Paper-
Curia-2Curia-2: Scaling Self-Supervised Learning for Radiology Foundation Modelsarxiv2026.02Paper-
OCTCube-MA 3D multimodal optical coherence tomography foundation model for retinal and systemic diseases with cross-cohort and cross-device validationNature Biomedical Engineering2026.04PaperProject
NeuroSTORMTowards a general-purpose foundation model for fMRI analysisNBME2026.03PaperProject
TriadVision foundation model for 3D magnetic resonance imaging segmentation, classification, and registrationMedIA2026.05Paper-
Foundation-VAEFoundation VAEs for 3D CT Reconstruction, Augmentation, and GenerationICML 20262026.05PaperProject
CoralBayCoralBay: A Self-Supervised CT Foundation ModelarXiv2026.06Paper-
How Much MRI Preprocessing Is Enough? A Cost-Utility Study for Brain MRI Foundation ModelsarXiv2026.06Paper-
NeuroVFMHealth system learning enables generalist neuroimaging modelsNature Medicine2026.07PaperProject
BrainFIBREBrainFIBRE: A Foundation Model via Information Decomposition for Brain MicrostructureECCV 2026 / arXiv2026.07Paper-
BoneCoT / BoneFMBoneCoT: multicentre validation of a whole-body skeleton foundation model for bone metastases guided by clinician-derived chain of thoughtNature Biomedical Engineering2026.07PaperProject
Cardiac CT FMA Unified Framework for Comprehensive Cardiac CT Segmentation and Phenotyping: Human-in-the-Loop Data Annotation, Vision Foundation Model Development, Multicenter Evaluation and Clinical ValidationarXiv2026.07Paper-
COJEPAContrastive Joint-Embedding Prediction for Representation Learning in Structural MRIarXiv2026.07Paper-
BrainNextBrainNext: A General-Purpose Self-Supervised Foundation Model for Brain MRI AnalysisarXiv2026.07Paper-
OrganLensOrganLens: Organ-Specific Representation Learning for CT Foundation ModelsarXiv2026.07Paper-
Rad-JEPA 3DRad-JEPA 3D: Radiology Joint-Embedding Predictive Model for 3D Computed TomographyarXiv2026.07Paper-

Contrastive Learning (CLIP-based Methods)

MethodDimensionTitleVenueDatePaperProject
MedCLIP2DMedCLIP: Contrastive Learning from Unpaired Medical Images and TextEMNLP 20222022.12PaperProject
PubMedCLIP2DPubMedCLIP: How Much Does CLIP Benefit Visual Question Answering in the Medical Domain?EACL 2023 (Findings)2023.05PaperProject
MedBLIP2DMedBLIP: Bootstrapping Language-Image Pre-training from 3D Medical Images and Textsarxiv2023.05Paper-
CLIP-Lung2DCLIP-Lung: Textual Knowledge-Guided Lung Nodule Malignancy PredictionMICCAI 20232023.10PaperProject
BioMedCLIP2DBiomedCLIP: A Multimodal Biomedical Foundation Model Pretrained from Fifteen Million Scientific Image-Text PairsNEJM AI 20242024.01PaperProject
PMC-CLIP2DPMC-CLIP: Contrastive Language-Image Pre-training using Biomedical DocumentsMICCAI 20232023.10PaperProject
UniMedCLIP2DUniMed-CLIP: Towards a Unified Image-Text Pretraining Paradigm for Diverse Medical Imaging ModalitiesArxiv2024.12PaperProject
ConceptCLIP2DConceptCLIP: Towards Trustworthy Medical AI via Concept-Enhanced Contrastive Language-Image Pre-trainingarXiv2025.01PaperProject
MMKD-CLIP2DUnifying Biomedical Vision-Language Expertise: Towards a Generalist Foundation Model via Multi-CLIP Knowledge DistillationArxiv2025.06Paper-
RadiSimCLIP2DRadiSimCLIP: A Radiology Vision-Language Model Pretrained on Simulated Radiologist Learning Dataset for Zero-Shot Medical Image UnderstandingMICCAI 2025 Workshop2025.10PaperProject
UniBrain3DUniBrain: Universal Brain MRI Diagnosis with Hierarchical Knowledge-enhanced Pre-trainingComputerized Medical Imaging and Graphics2023.09PaperProject
CT2Rep3DCT2Rep: Automated Radiology Report Generation for 3D Medical ImagingMICCAI 20242024.03PaperProject
CT-CLIP3DGeneralist Foundation Models from a Multimodal Dataset for 3D Computed TomographyNat Biomed Eng2024.03PaperProject
CT-GLIP3DCT-GLIP: 3D Grounded Language-Image Pretraining with CT Scans and Radiology Reports for Full-Body ScenariosarXiv2024.04Paper-
RadCLIP3DRadCLIP: Enhancing Radiologic Image Analysis Through Contrastive Language–Image PretrainingTNNLS2024.03Paper-
Percival3DA Pan-Organ Vision-Language Model for Generalizable 3D CT RepresentationsmedRxiv2025.07Paper-
OpenVocabCT3DTowards Universal Text-driven CT Image SegmentationarXiv2025.03Paper-
fVLM3DLarge-scale and Fine-grained Vision-language Pre-training for Enhanced CT Image UnderstandingICLR 20252025.03PaperProject
HLIP3DTowards Scalable Language-Image Pre-training for 3D Medical Imagingarxiv2025.05PaperProject
RadZero3D3DRadZero3D: Bridging Self-Supervised Video Models and Medical Vision-Language Alignment for Zero-Shot Chest CT InterpretationICCV 2025 Workshop2025.10Paper-
T3D3DT3D: Advancing 3D Medical Vision-Language Pre-training by Learning Multi-View Visual ConsistencyICCV 2025 Workshop2025.10Paper-
ViSD-Boost3DBoosting Vision Semantic Density with Anatomy Normality Modeling for Medical Vision-language Pre-trainingICCV 20252025.08PaperProject
VELVET-Med3DVELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in MedicinearXiv2025.08Paper-
MedVista3D3DMedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and ReportingarXiv2025.09Paper-
COLIPRI3DComprehensive Language-Image Pre-training for 3D Medical Image UnderstandingarXiv2025.10PaperProject
MPS-CT3DMore performant and scalable: Rethinking contrastive vision-language pre-training of radiology in the LLM eraMICCAI 20252025.10PaperProject
PET-CLIP Captioner3DLocation-Guided Automated Lesion Captioning in Whole-body PET/CT ImagesMICCAI 20252025.10Paper-
MR-CLIP3DMetadata-Aligned 3D MRI Representations for Contrast Understanding and Quality ControlarXiv2025.11Paper-
BrgSA3DBridged Semantic Alignment for Zero-shot 3D Medical Image DiagnosisarXiv2025.11PaperProject
SCALE-VLP3DSCALE-VLP: Soft-Weighted Contrastive Volumetric Vision-Language Pre-training with Spatial-Knowledge SemanticsarXiv2025.11Paper-
SPECTRE3DScaling Self-Supervised and Cross-Modal Pretraining for Volumetric CT TransformersCVPR 20262025.11PaperProject
Pillar-03DPillar-0: A New Frontier for Radiology Foundation Modelsarxiv2025.11PaperProject
BTB3D3DBetter Tokens for Better 3D: Advancing Vision-Language Modeling in 3D Medical ImagingNeurIPS 20252025.12PaperProject
NeuroVFM3DNeuroVFM: A Contrastive Vision-Language Model for Medical Reasoning in Alzheimer's Disease DiagnosisWACV 2026 Workshops2026.01Paper-
TotalFM3DTotalFM: An Organ-Separated Framework for 3D-CT Vision Foundation Modelsarxiv2026.01Paper-
MG-3D3DMG-3D: Multi-Grained Knowledge-Enhanced 3D Medical Vision-Language Pre-trainingMedical Image Analysis (MedIA)2026.01PaperProject
MedMAP3D3D Modality-Aware Pre-training for Vision-Language Model in MRI Multi-organ Abnormality Detectionarxiv2026.02PaperProject
Prima3DLearning neuroimaging models from health system-scale dataNature Biomedical Engineering2026.02PaperProject
SigVLP3DSigVLP: Sigmoid Volume-Language Pre-Training for Self-Supervised CT-Volume Adaptive Representation LearningarXiv2026.02Paper-
RadFinder3DLearning to Read Where to Look: Disease-Aware Vision–Language Pretraining for 3D CTarXiv2026.03PaperProject
Decipher-MR3DDecipher-MR: A Vision-Language Foundation Model for 3D MRI Representationsnpj Digital Medicine2026.04Paper-
3DCLIP Architecture for Abdominal CT Image–Text Alignment and Zero-Shot Learning: Investigating Batch Composition and Data Scalingarxiv2026.04Paper-
ASAP3DASAP: Advancing Medical Volumetric Representation Learning with Anatomy-aware Semantically-adaptive Pre-trainingarXiv2026.05Paper-
GLeVE3DGLeVE: Graph-Guided Lesion Grounding with Proposal Verification in 3D CTarXiv2026.05Paper-
CA-GCL3DCA-GCL: Cross-Anatomy Global-Local Contrastive Learning for Robust 3D Medical Image UnderstandingarXiv2026.05Paper-
SegReg-Rep3DSegReg-Rep: Region-Aware Vision-Language Alignment for Fine-Grained Radiology Report Generation from 3D Medical ImagesIEEE TPRMS2026.05PaperProject
GLINT3DGLINT: Sparsely Gated Vision-Language Alignment for Fine-Grained Radiology RepresentationsarXiv2026.06Paper-
RadGrounder2DScalable Training of Spatially Grounded 2D Vision-Language Models for RadiologyarXiv2026.06Paper-
RenalCLIP3DA Disease-Centric Vision-Language Foundation Model for Precision Oncology in Kidney CancerNature Communications2026.06Paper-
Jolia / ConQuer3DJolia: Concept-Level Vision-Language Alignment for 3D CT Contrastive LearningarXiv2026.06Paper-
3DDisease-Centric Vision-Language Pretraining with Hybrid Visual Encoding for 3D Computed TomographyarXiv2026.06Paper-
MedReCo3DA Vision-language Framework for Comparative Reasoning in RadiologyarXiv2026.06Paper-
SuG3DSuper-Generalist: Towards Comprehensive and Accurate Medical Image Understanding via Generalist-Specialist SynergyarXiv2026.07Paper-
OKA-CT3DLearning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report KnowledgearXiv2026.07Paper-
OCP-CT3DFine-Grained Vision-Language Pretraining with Organ-Conditioned Pattern Tokens for CT UnderstandingarXiv2026.07Paper-
MseaCL3DMultimodal Semantic-Aware Contrastive Learning For False Negative Mitigation in 3D Medical ImagingarXiv2026.07Paper-
CARVE3DWhen Can Test-Time Adaptation Help Zero-Shot CT Vision-Language Models?arXiv2026.07Paper-
ACA3DAnatomy Contextualized Adaption of CT Foundation ModelsarXiv2026.07Paper-
Spectrum3DLearning How Much, Not Just What: Cross-Patient Burden Order for CT Vision-Language PretrainingarXiv2026.07Paper-
SCOPE3DSemantically Calibrated Evidence Composition for CT Vision-Language LearningarXiv2026.07Paper-

MLLM-based Methods

2D-only

MethodTitleVenueDatePaperProject
BiomedGPTBiomedGPT: A Generalist Vision-Language Foundation Model for Diverse Biomedical TasksNat Med 20242023.05PaperProject
MedVInTPMC-VQA: Visual Instruction Tuning for Medical Visual Question AnsweringArxiv/ Communications Medicine2023.05/2024.12Paper-
LLaVA-MedLLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One DayNeurIPS 20232023.06PaperProject
Med-FlamingoMed-Flamingo: a Multimodal Medical Few-shot LearnerML4H 20232023.07PaperProject
Med-PaLM MTowards Generalist Biomedical AINEJM AI 2024 / arXiv2023.07PaperProject
Qilin-Med-VLQilin-Med-VL: Towards Chinese Large Vision-Language Model for General HealthcarearXiv2023.10Paper-
R2GenGPTR2GenGPT: Radiology Report Generation with Frozen LLMsMeta-Radiology2023.11Paper-
BiRDA Refer-and-Ground Multimodal Large Language Model for BiomedicineMICCAI 20242024.06Paper-
HuatuoGPT-VisionTowards Injecting Medical Visual Knowledge into Multimodal LLMs at ScaleEMNLP242024.06PaperProject
Llama3-MedAdvancing High Resolution Vision-Language Models in BiomedicinearXiv2024.06PaperProject
MiniGPT-MedMiniGPT-Med: Large Language Model as a General Interface for Radiology DiagnosisarXiv2024.07PaperProject
TinyLLaVA-MedDemocratizing MLLMs in Healthcare: TinyLLaVA-Med for Efficient Healthcare Diagnostics in Resource-Constrained SettingsMICCAI242024.09Paper-
Med-MoEMed-MoE: Mixture of Domain-Specific Experts for Lightweight Medical Vision-Language ModelsEMNLP242024.09PaperProject
GMAI-VLGMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AIarXiv/AAAI262024.11PaperProject
BiMediX2BiMediX2: Bio-Medical EXpert LMM for Diverse Medical ModalitiesFindings of EMNLP 20252025.01PaperProject
HealthGPTHealthGPT: A Medical Large Vision-Language Model for Unifying Comprehension and Generation via Heterogeneous Knowledge AdaptationICML 20252025.02PaperProject
MedVLM-R1MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement LearningMICCAI252025.02PaperProject
Med-R1Med-R1: Reinforcement Learning for Generalizable Medical Reasoning in Vision-Language ModelsTMI2025.03PaperProject
OmniV-MedOmniV-Med: Scaling Medical Vision-Language Model for Universal Visual UnderstandingarXiv2025.04Paper-
UniBioMedUniBiomed: A Universal Foundation Model for Grounded Biomedical Image InterpretationarXiv2025.04PaperProject
MedRegAMedRegA: Interpretable Bilingual Multimodal Large Language Model for Diverse Biomedical TasksICLR252025.04PaperProject
UMed-LVLMImproving Medical Large Vision-Language Models with Abnormal-Aware FeedbackACL 20252025.05Paper-
QoQ-MedQoQ-Med: Building Multimodal Clinical Foundation Models with Domain-Aware GRPO TrainingarXiv2025.06PaperProject
LingshuLingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and ReasoningarXiv2025.06PaperProject
MedGemmaMedGemma Technical ReportarXiv2025.07PaperProject
Critus-VCitrus-V: Advancing Medical Foundation Models with Unified Medical Image Grounding for Clinical ReasoningarXiv2025.09PaperProject
MedPLIBTowards a Multimodal Large Language Model with Pixel-Level Insight for BiomedicineAAAI252025.10PaperProject
OctoMedOctoMed: Data Recipes for State-of-the-Art Multimodal Medical ReasoningarXiv2025.11PaperProject
MedMOMedMO: Grounding and Understanding Multimodal Large Language Model for Medical ImagesarXiv2026.02PaperProject
MediX-R1MediX-R1: Open Ended Medical Reinforcement LearningarXiv2026.02PaperProject
MEDIC-ADMEDIC-AD: Towards Medical Vision-Language Model’s Clinical IntelligenceCVPR26 Oral2026.03PaperProject
MedVRMedVR: Annotation-Free Medical Visual Reasoning via Agentic Reinforcement LearningICLR 262026.04PaperProject

Unified 2D + Serialized-3D

MethodTitleVenueDatePaperProject
RadFMTowards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical DataNat Commun 2025 / arXiv2023.08PaperProject
Med-GeminiAdvancing Multimodal Medical Capabilities of GeminiNat Med 2025 / arXiv2024.05PaperProject
Med-2E3Med-2E3: A 2D-Enhanced 3D Medical Multimodal Large Language ModelBIBM252024.11Paper-
Hulu-MedHulu-Med: A Transparent Generalist Model towards Holistic Medical Vision-Language UnderstandingarXiv2025.10PaperProject
Fleming-VLFleming-VL: Towards Universal Medical Visual Reasoning with Multimodal LLMsarXiv2025.11PaperProject
MedM-VLMedM-VL: What Makes a Good Medical LVLM?International Workshop on Agentic AI for Medicine 20252025.09PaperProject
CTInstructCTInstruct: Towards Unified 3D CT Understanding via Instruction TuningAAAI 262026.01PaperProject
A data-efficient 3D medical vision-language model using only a 2D encoderScientific report2026.02Paper-
MedPrunerMedPruner: Training-Free Hierarchical Token Pruning for Efficient 3D Medical Image Understanding in Vision-Language ModelsMICCAI20262026.03Paper-
PhotonPhoton: Speedup Volume Understanding with Efficient Multimodal Large Language ModelsICLR 262026.03PaperProject
OmniCTOmniCT: Towards a Unified Slice-Volume LVLM for Comprehensive CT AnalysisICLR 262026.03PaperProject
MedGemma1.5MedGemma 1.5 Technical ReportarXiv2026.04PaperProject
TGH-MoEAdapting 2D Multi-Modal Large Language Model for 3D CT Image AnalysisarXiv2026.04Paper-
Brain-AdapterBrain-Adapter: A Dual-Stream Vision-Language MIL Framework for Comprehensive 3D CT Diagnosis of Acute Intracranial PathologiesarXiv2026.06Paper-
UniReason-MedUniReason-Med: A Shared Grounded Reasoning Interface for 2D-to-3D Transfer in Medical VQAarXiv2026.06PaperProject
MedReCo-VLMA Vision-language Framework for Comparative Reasoning in RadiologyarXiv2026.06Paper-
RadSightRadSight: Towards Perceptually Reliable Multimodal Radiology Image UnderstandingarXiv2026.07PaperProject
ClinFusionClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical UnderstandingarXiv2026.07PaperProject
HounsfieldTowards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language ModelsarXiv2026.07PaperProject
MedARCMedARC: Training-Free Adaptive Redundancy Compression of Visual Tokens for 3D Medical Vision-Language ModelsarXiv2026.07Paper-
ORCAORCA: ORgan-Centroid Aggregation for Training-Free 3D CT Visual Token CompressionarXiv2026.07PaperProject

Native 3D-Volume

MethodTitleVenueDatePaperProject
CT2RepCT2Rep: Automated Radiology Report Generation for 3D Medical ImagingMICCAI 2024 / arXiv2024.03PaperProject
Dia-LLaMADia-LLaMA: Towards Large Language Model-driven CT Report GenerationMICCAI2024.03Paper-
M3D-LaMedM3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language ModelsICLR 2025 / arXiv2024.04PaperProject
MerlinMerlin: A Computed Tomography Vision-Language Foundation Model and DatasetNature 2026 / arXiv2024.06PaperProject
BrainGPTTowards a Holistic Framework for Multimodal Large Language Models in Three-dimensional Brain CT Report GenerationNat Commun 2025 / arXiv2024.07PaperProject
3D-CT-GPT3D-CT-GPT: Generating 3D Radiology Reports through Integration of Large Vision-Language ModelsarXiv2024.09Paper-
E3D-GPTE3D-GPT: Enhanced 3D Visual Foundation for Medical Vision-Language ModelarXiv2024.10Paper-
Reg2RGLarge Language Model with Region-guided Referring and Grounding for CT Report GenerationarXiv2024.11Paper-
MS-VLMRead Like a Radiologist: Efficient Vision-Language Model for 3D Medical Imaging InterpretationarXiv2024.12Paper-
MEPNetMEPNet: Medical Entity-balanced Prompting Network for Brain CT Report GenerationarXiv2025.03Paper-
Med3DVLMMed3DVLM: An Efficient Vision-Language Model for 3D Medical Image AnalysisIEEE JBHI 2025 / arXiv2025.03PaperProject
HSENetHSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language UnderstandingarXiv2025.06Paper-
MedRegion-CTMedRegion-CT: Region-Focused Multimodal LLM for Comprehensive 3D CT Report GenerationarXiv2025.06Paper-
mpLLMMultimodal LLM With Hierarchical Mixture-of-Experts for VQA on 3D Brain MRIarXiv2025.09Paper-
3DReasonKnee3DReasonKnee: Advancing Grounded Reasoning in Medical Vision Language ModelsarXiv2025.10Paper-
PETARPETAR: Localized Findings Generation with Mask-Aware Vision-Language Modeling for PET Automated ReportingarXiv2025.10Paper-
BTB3DBetter Tokens for Better 3D: Advancing Vision-Language Modeling in 3D Medical ImagingNeurlPS 20252025.10PaperProject
PETRG-3DVision-Language Models for Automated 3D PET/CT Report GenerationarXiv2025.11Paper-
CTest-MetricCTest-Metric: A Unified Framework to Assess Clinical Validity of Metrics for CT Report GenerationISBI20262026.01Paper-
Brain3DBrain3D: Brain Report Automation via Inflated Vision Transformers in 3DarXiv2026.02PaperProject
Med3D-R1Med3D-R1: Incentivizing Clinical Reasoning in 3D Medical Vision-Language Models for Abnormality DiagnosisarXiv2026.02Paper-
LoV3DLoV3D: Grounding Cognitive Prognosis Reasoning in Longitudinal 3D Brain MRI via Regional Volume AssessmentsarXiv2026.03Paper-
Ker-VLJEPA-3BCurriculum-Driven 3D CT Report Generation via Language-Free Visual Grafting and Zone-Constrained CompressionarXiv2026.03PaperProject
U-VLMU-VLM: Hierarchical Vision Language Modeling for Report GenerationarXiv2026.02PaperProject
CT-CHATGeneralist foundation models from a multimodal dataset for 3D computed tomographyNature Biomedical Engineering2026.02PaperProject
MedVL-SAM2MedVL-SAM2: A unified 3D medical vision–language model for multimodal reasoning and prompt-driven segmentationarXiv2026.01Paper-
BoiDEnhancing 3D medical multi-modal large language models with integrated human body priors for computed tomographyPattern Recognition2026.04Paper-
DCP-PDEnhancing Fine-Grained Spatial Grounding in 3D CT Report Generation via Discriminative Guidancearxiv2026.04Paper-
SegReg-RepSegReg-Rep: Region-Aware Vision-Language Alignment for Fine-Grained Radiology Report Generation from 3D Medical ImagesIEEE TPRMS2026.05PaperProject
CLarGenGenerating Reports or Repeating Templates? Measuring and Mitigating Template Collapse in 3D CT Report GenerationarXiv2026.05Paper-
TIF-GRPORegulating Anatomy-Aware Rewards via Trajectory-Integral Feedback for Volumetric Computed Tomography AnalysisarXiv2026.05Paper-
RAD3D-PrefixRevisiting LLM Adaptation for 3D CT Report Generation: A Study of Scaling and Diagnostic PriorsarXiv2026.06Paper-
E-MRLE-MRL: Cross-view Aligned Evidence-driven Multimodal Reinforcement Learning for Reliable 3D Tumor AnalysisarXiv2026.06Paper-
MRI2RepMRI2Rep: Autoregressive Structured Report Generation for 3D Liver MRIarXiv2026.06Paper-
NeuroVFMHealth system learning enables generalist neuroimaging modelsNature Medicine2026.07PaperProject
PIPAPIPA: Prior-Driven Prompting with Diagnosis-Oriented Retrieval-Augmentation for 3D Radiology Report GenerationIEEE TMI2026.07PaperProject
MonteRETMonteRET: AI Agent Enhancing Multimodal LLMs with Multi-granularity Knowledge Retrieval for Chest CT Report GenerationarXiv2026.07Paper-
Multi-LLM MRIMulti-LLM Collaborative MRI Report Generation for Visual Instruction Tuning in Brain OncologyarXiv2026.07Paper-

Agentic Systems

Core modules of an agentic workflow for 3D radiology analysis

Core modules of agentic 3D radiology workflows: reasoning and planning, tool-augmented perception, memory and context, and workflow collaboration.

MethodTitleVenueDatePaperProject
MDAgentsMDAgents: An Adaptive Collaboration of LLMs for Medical Decision-MakingNeurIPS 20242024.12 (arXiv: 2024.04)PaperProject
MMedAgentMMedAgent: Learning to Use Medical Tools with Multi-modal AgentFindings of EMNLP 20242024.11 (arXiv: 2024.07)PaperProject
MedAgent-ProMedAgent-Pro: Towards Evidence-based Multi-modal Medical Diagnosis via Reasoning Agentic WorkflowICLR 20262025.03PaperProject
DOLAAutonomous Radiotherapy Treatment Planning Using DOLA: A Privacy-Preserving, LLM-Based Optimization AgentarXiv2025.03Paper-
GPT-PlanA Feasibility Study of Automating Radiotherapy Planning with Large Language Model AgentsPhysics in Medicine and Biology2025.03Paper-
VILA-M3VILA-M3: Enhancing Vision-Language Models with Medical Expert KnowledgeCVPR252024.11PaperProject
CT-AgentCT-Agent: A Multimodal-LLM Agent for 3D CT Radiology Question AnsweringarXiv2025.05Paper-
M^3BuilderM^3Builder: A Multi-Agent System for Automated Machine Learning in Medical ImagingAI for Clinical Applications 20252025.05Paper-
SAMIRATowards user-centered interactive medical image segmentation in VR with an assistive AI agentarXiv2025.05Paper-
MAMMAM: Modular Multi-Agent Framework for Multi-Modal Medical Diagnosis via Role-Specialized CollaborationFindings of ACL 20252025.07 (arXiv: 2025.06)PaperProject
AgentMRIAgentMRI: A Vision Language Model-Powered AI System for Self-regulating MRI Reconstruction with Multiple DegradationsJournal of Imaging Informatics in Medicine2025.07Paper-
CTPA-AgentVision-language model for report generation and outcome prediction in CT pulmonary angiogramnpj Digital Medicine2025.07PaperProject
TissueLabA co-evolving agentic AI system for medical imaging analysisarXiv2025.09PaperProject
Scan-do AttitudeScan-do Attitude: Towards Autonomous CT Protocol Management Using a Large Language Model AgentAgentic AI for Medicine / Springer2025.09Paper-
VoxelPromptVoxelPrompt: A Vision Agent for End-to-End Medical Image AnalysisarXiv2025.10Paper-
MedAgentSimMedAgentSim: Self-Evolving Multi-Agent Simulations for Realistic Clinical InteractionsMICCAI 20252025.10 (arXiv: 2025.03)PaperProject
AURAAURA: A Multi-Modal Medical Agent for Understanding, Reasoning & AnnotationMICCAI Workshop 20252025.10 (arXiv: 2025.07)PaperProject
MedEyesMedEyes: Learning Dynamic Visual Focus for Medical Progressive DiagnosisarXiv2025.11PaperProject
MedSAM3MedSAM3: Delving into Segment Anything with Medical ConceptsarXiv2025.11PaperProject
Radiologist CopilotRadiologist Copilot: An Agentic Framework Orchestrating Specialized Tools for Reliable Radiology ReportingarXiv2025.12Paper-
INFORM-CTINFORM-CT: INtegrating LLMs and VLMs FOR Incidental Findings Management in Abdominal CTMIDL 20262025.12PaperProject
IBISAgentIBISAgent: Reinforcing Pixel-Level Visual Reasoning in MLLMs for Universal Biomedical Object Referring and SegmentationarXiv2026.01Paper-
MedVistaGymMEDVISTAGYM: A Scalable Training Environment for Thinking with Medical Images via Tool-Integrated Reinforcement LearningarXiv2026.01Paper-
An Explainable Agentic AI Framework for Uncertainty-Aware and Abstention-Enabled Acute Ischemic Stroke Imaging DecisionsarXiv2026.01Paper-
LungNoduleAgentLungNoduleAgent: A Collaborative Multi-Agent System for Precision Diagnosis of Lung NodulesAAAI 20262026.02 (arXiv: 2025.11)PaperProject
3DMedAgent3DMedAgent: Unified Perception-to-Understanding for 3D Medical AnalysisarXiv2026.02PaperProject
MedSAM-AgentMedSAM-Agent: Empowering Interactive Medical Image Segmentation with Multi-turn Agentic Reinforcement LearningarXiv2026.02PaperProject
CARECARE: Towards Clinical Accountability in Multi-Modal Medical Reasoning with an Evidence-Grounded Agentic FrameworkICLR 20262026.05 (arXiv: 2026.03)PaperProject
ToolSelectPicking the Right Specialist: Attentive Neural Process-based Selection of Task-Specialized Models as Tools for Agentic Healthcare SystemsarXiv2026.02Paper-
CoMMaCoMMa: Contribution-Aware Medical Multi-Agents From A Game-Theoretic PerspectivearXiv2026.02Paper-
MedSegAgentMedSegAgent: A Universal and Scalable Multi-Agent System for Instructive Medical Image SegmentationIEEE JBHI2026.03PaperProject
CT-FlowCT-Flow: Orchestrating CT Interpretation Workflow with Model Context Protocol ServersarXiv2026.03Paper-
Agent-MIRAAgent-MIRA: AI-orchestrated Medical Imaging Agent for PET Image Retrieval and AssistanceComputerized Medical Imaging and Graphics2026.03Paper-
MeissaMeissa: Multi-modal Medical Agentic IntelligencearXiv2026.03PaperProject
BT-RADS AgentAgentic Automation of BT-RADS Scoring: End-to-End Multi-Agent System for Standardized Brain Tumor Follow-up AssessmentarXiv2026.03Paper-
TheraAgentTheraAgent: Multi-Agent Framework with Self-Evolving Memory and Evidence-Calibrated Reasoning for PET TheranosticsarXiv2026.03Paper-
MedOpenClawMEDOPENCLAW: Auditable Medical Imaging Agents Reasoning over Uncurated Full StudiesarXiv2026.03Paper-
MedMASLabMedMASLab: A Unified Orchestration Framework for Benchmarking Multimodal Medical Multi-Agent SystemsarXiv2026.03PaperProject
ClinicalAgentsClinicalAgents: Multi-Agent Orchestration for Clinical Decision Making with Dual-MemoryarXiv2026.03Paper-
Doctorina MedBenchDoctorina MedBench: End-to-End Evaluation of Agent-Based Medical AIarXiv2026.03Paper-
SEERSkill-Evolving Grounded Reasoning for Free-Text Promptable 3D Medical Image SegmentationarXiv2026.03Paper
RadAgentRadAgent: A Tool-Using AI Agent for Stepwise Interpretation of Chest Computed TomographyarXiv2026.04PaperProject
BAAI Cardiac AgentBAAI Cardiac Agent: An intelligent multimodal agent for automated reasoning and diagnosis of cardiovascular diseases from cardiac magnetic resonance imagingarXiv2026.04PaperProject
DosimeTronDosimeTron: Automating Personalized Monte Carlo Radiation Dosimetry in PET/CT with Agentic AIarXiv2026.04Paper-
MARCHMARCH: Multi-Agent Radiology Clinical Hierarchy for CT Report GenerationACL 20262026.04Paper-
Neuro-Radiological AgentAgentic Large Language Models for Training-Free Neuro-Radiological Image AnalysisarXiv2026.04Paper-
Agent4MRAgentic MR sequence development: leveraging LLMs with MR skills for automatic physics-informed sequence developmentarXiv2026.04Paper-
Artifact-based Agent FrameworkAn Artifact-based Agent Framework for Adaptive and Reproducible Medical Image ProcessingarXiv2026.04Paper-
NeuroClawNeuroClaw: Closed-Loop Agentic AI for Executable and Reproducible Neuroimaging ResearcharXiv2026.04Paper-
Neuro-OracleNeuro-Oracle: A Trajectory-Aware Agentic RAG Framework for Interpretable Epilepsy Surgical PrognosisarXiv2026.04Paper-
MedScribeMedScribe: Clinically Grounded CT Reporting through Agentic WorkflowsarXiv2026.05Paper-
GAZEGAZE: Grounded Agentic Zero-shot Evaluation with Viewer-Level Tools and Literature Retrieval on Rare Brain MRIarXiv2026.05Paper-
NeuroAgentNeuroAgent: LLM Agents for Multimodal Neuroimaging Analysis and ResearcharXiv2026.05Paper-
NEXUSTowards a Virtual Neuroscientist: Autonomous Neuroimaging Analysis via Multi-Agent CollaborationarXiv2026.05PaperProject
M2M-LLM-RTA Machine-to-Machine Knowledge-Guided LLM Agent for Generalizable Radiotherapy Treatment PlanningarXiv2026.05Paper-
SpineAgentA Multi-Agent System for Spine MRI Report Generation from Multi-Sequence ImagingarXiv2026.06Paper-
MedToolicaMedToolica: Finetuning-Free Agentic Compositional Tool Learning for 3D CT ReasoningMachine Learning and Knowledge Extraction2026.06PaperProject
MARTPMARTP: A Multi-Agent Simulation Framework for Automated Radiation Therapy Planning Based on LLMsPhysics in Medicine and Biology2026.06Paper-
SAGEAutomated Stereotactic Radiosurgery Planning Using a Human-in-the-Loop Reasoning Large Language Model AgentResearch Square2026.06Paper-
PET/CT AgentEnd-to-End PET/CT Interpretation and Quantification with an LLM-Orchestrated AI Agent: A Real-World Pilot StudyJournal of Nuclear Medicine2026.06Paper-
PD-CTAgentPolicy-Driven CT-Agent: Modeling Phase-Aware Diagnostic Control for Clinically Consistent CT ReasoningarXiv2026.07Paper-
MonteRETMonteRET: AI Agent Enhancing Multimodal LLMs with Multi-granularity Knowledge Retrieval for Chest CT Report GenerationarXiv2026.07Paper-
One-for-AllOne-for-All Adaptive Radiotherapy Planning Agent: A Foundation Framework for Daily CBCT-guided RadiotherapyarXiv2026.07Paper-

Datasets & Benchmarks

DatasetTitleDateVenuePaper LinkProject
SLIVER07Segmentation in the liver 2007 (SLIVER07) challenge2007.09MICCAI WorkshopPaperProject
SKI10Segmentation of Knee Images 2010 (SKI10)2010.09MICCAI ChallengePaperProject
LIDC-IDRIData From LIDC-IDRI: The Lung Image Database Consortium and Image Database Resource Initiative2011.06Med PhysPaperProject
LOLA11LOLA11: LObe and Lung Analysis 2011 Challenge2011.09MICCAI WorkshopProject
STACOM 2011 Motion TrackingSTACOM 2011: Cardiac Motion Tracking Challenge2011.09MICCAI WorkshopPaperProject
Mindboggle-101Mindboggle-101: Evaluating Brain Image Labeling Methods2012.09NeuroImagePaperProject
PROMISE12PROMISE12: Prostate MR Image Segmentation 2012 Challenge2012.10MICCAI ChallengePaperProject
NLSTThe National Lung Screening Trial: overview and study design2013.01RadiologyProject
Farsiu Ophthalmology 2013Quantitative Classification of Eyes with and without Intermediate Age-related Macular Degeneration Using Optical Coherence Tomography2013.03OphthalmologyPaperProject
Prostate-3TData From Prostate-3T2013.06TCIA CollectionProject
MRBrainS13MRBrainS13: Grand Challenge on MR Brain Image Segmentation2013.09MICCAI ChallengeProject
Chiu BOE 2014Kernel regression based segmentation of optical coherence tomography images with diabetic macular edema2014.01Biomed Opt ExpressPaperProject
Srinivasan BOE 2014Fully automated detection of diabetic macular edema and dry age-related macular degeneration from optical coherence tomography images2014.03Biomed Opt ExpressPaperProject
orCaScoreAn evaluation of automatic coronary artery calcium scoring methods with cardiac CT using the orCaScore framework2014.09MICCAI ChallengePaperProject
CETUS2014CETUS: Cardiac Echocardiography Tracking and Segmentation Challenge2014.09MICCAI ChallengeProject
Prostate-DiagnosisPROSTATE-DIAGNOSIS: Multiparametric MRI for Prostate Cancer2015.03TCIA CollectionProject
BTCVMICCAI multi-atlas labeling beyond the cranial vault-workshop and challenge2015.04MICCAI WorkshopProject
ISMRM2015 HARDIISMRM 2015 Tractography Challenge2015.06ISMRM ChallengeProject
NEATBrainS15NEATBrainS15: Neonatal Brain Structure Segmentation2015.09MICCAI ChallengeProject
PDDCAPublic Domain Database for Computational Anatomy: Head and Neck2015.09MICCAI ChallengeProject
HVSMR 2016HVSMR 2016: Whole Heart and Great Vessel Segmentation Challenge2016.07MICCAI ChallengePaperProject
MSSEG 2016Objective Evaluation of Multiple Sclerosis Lesion Segmentation using a Data Management and Processing Infrastructure2016.10MICCAI Challenge/NaturePaperProject
PROSTATExPROSTATEx: PROSTATE MR Image Dataset With Prostate Cancer Annotations2016.10SPIE-AAPM-NCI PROSTATEx ChallengeProject
WMHWMH Segmentation Challenge: White Matter Hyperintensity Segmentation in Brain MR2017.03MICCAI ChallengePaperProject
LGG-1p19qDeletionLGG-1p19qDeletion: Low-Grade Glioma MRI with Genomic Annotations2017.03TCIA CollectionProject
PROSTATEx-2PROSTATEx-2: Lesion Classification Challenge2017.06AAPM Grand ChallengeProject
ACDCAutomatic Cardiac Diagnosis Challenge2017.09STACOM / MICCAIPaperProject
RETOUCHRETOUCH: Retinal OCT Fluid Segmentation Challenge2017.09MICCAI ChallengePaperProject
ROCCROCC: Retinal OCT Classification Challenge2017.09MICCAI ChallengeProject
iSeg2017iSeg-2017: Infant Brain MRI Segmentation Challenge2017.09MICCAI ChallengePaperProject
DeepLesionDeepLesion: automated mining of large-scale lesion annotations and universal lesion detection in CT2017.10JMI / arXivPaperProject
LUNA 16Validation, comparison, and combination of algorithms for automatic detection of pulmonary nodules in computed tomography images: the LUNA 16 challenge2017.12Medical Image AnalysisPaperProject
Mandibular-CT-DatasetMandibular CT Dataset Collection for 3D Reconstruction and Segmentation2018.03figsharePaperProject
FUMPEComputer-aided detection of pulmonary embolism in CT2018.03arXiv / KaggleProject
ISLES 2018ISLES 2018 – Ischemic Stroke Lesion Segmentation2018.09MICCAI ChallengePaperProject
MRBrainS18MRBrainS18: MR Brain Segmentation Challenge 20182018.09MICCAI ChallengeProject
Atrial Segmentation Challenge2018 Atrial Segmentation Challenge2018.09MICCAI ChallengeProject
IVDM3SegIVDM3Seg: Intervertebral Disc and Vertebrae Segmentation Challenge2018.09MICCAI ChallengeProject
MRNetMRNet: Knee MRI Dataset for Abnormality Detection2018.09NIPS WorkshopProject
OCT Glaucoma DetectionGlaucoma Detection in 3D Spectral-Domain OCT2018.10Sci RepProject
BraTSBrain Tumor Segmentation (BraTS) Challenge2018.11MICCAI Challenge (series)PaperProject
fastMRIfastMRI: A Publicly Available Raw k-Space and DICOM Dataset of Knee and Brain MR Images2018.11MRMPaperProject
LiTSLiver Tumor Segmentation (LiTS) Challenge2019.01MICCAI ChallengePaperProject
OASIS-3OASIS-3: Longitudinal Neuroimaging, Clinical, and Cognitive Dataset for Normal Aging and Alzheimer Disease2019.01Sci DataPaperProject
MM-WHSMM-WHS: Multi-Modality Whole Heart Segmentation2019.02MICCAI ChallengePaperProject
CHAOS CT-MRICHAOS - Combined (CT-MR) Healthy Abdominal Organ Segmentation2019.02ISBI ChallengePaperProject
KiTS19KiTS19: Kidney Tumor Segmentation Challenge2019.04MICCAI ChallengePaperProject
AAPM-RT-MACAAPM RT-MAC: MR-only based Radiotherapy in Head and Neck2019.07AAPM ChallengePaperProject
iSeg-2019iSeg-2019: Infant Brain MRI Segmentation Challenge2019.09MICCAI ChallengePaperProject
SegTHORSegTHOR: Segmentation of thoracic organs at risk in CT images2019.09Physica MedicaPaperProject
VerSe20VerSe 2020: Vertebral Segmentation Challenge at MICCAI2020.01MICCAI ChallengeProject
VerSe19VerSe 2019: Vertebral Segmentation Challenge at MICCAI2020.01MICCAI ChallengePaperProject
COVID-19-CT-SegCOVID-19 CT lung and infection segmentation dataset2020.04zenodoProject
M&MsM&Ms: Multi-Centre, Multi-Vendor & Multi-Disease Cardiac MR Segmentation Challenge2020.05MICCAI ChallengeProject
CTPelvic1KCTPelvic1K: A Large-Scale Pelvic CT Dataset for Multi-Task Parsing2020.06arXivPaperProject
Prostate MR Segmentation Dataset (SAML)Federated Domain Generalization on Medical Image Segmentation via Episodic Learning in Continuous Frequency Space2020.09MICCAIProject
EMIDECEMIDEC 2020: Myocardial Infarction Detection, Segmentation and Classification2020.09MICCAI ChallengePaperProject
KNOAP2020KNOAP2020: Knee Osteoarthritis Progression Prediction Challenge2020.09MICCAI ChallengeProject
Learn2Reg Lung CTLearn2Reg 2020: Lung CT Registration2020.09MICCAI ChallengeProject
Learn2Reg Abdomen CT-CTLearn2Reg 2020: Abdominal CT-CT Registration2020.09MICCAI ChallengeProject
RibFrac2020RibFrac: Rib Fracture Detection and Classification Challenge2020.10MICCAI ChallengeProject
HECKTOR 2020HECKTOR 2020: Segmentation of Head and Neck Tumor in PET/CT2020.11MICCAI ChallengeProject
CT-ORGCT-ORG, a new dataset for multiple organ segmentation in computed tomography2020.11NaturePaperProject
RAD-ChestCTMachine-Learning-Based Multiple Abnormality Prediction with Large-Scale Chest Computed Tomography Volumes2021.01Medical Image AnalysisPaperProject
Eye OCT Datasets (3D)3D Retinal OCT Classification and Segmentation Dataset2021.01TianchiProject
HECKTOR 2021HECKTOR 2021: Head and Neck Tumor Segmentation and Outcome Prediction2021.05MICCAI ChallengeProject
CTSpine1KCTSpine1K: A Large-Scale Dataset for Spine Parsing in CT2021.07arXivProject
MSSEG-2MSSEG-2 challenge: Multiple Sclerosis Lesion Segmentation at 7T and 3T MRI2021.07NeuroImage ClinProject
FLARE21FLARE 2021: A Challenge on Abdominal Multi-organ Segmentation2021.09MICCAI ChallengeProject
QUBIQ2021 3D CTQUBIQ 2021: Quantification of Uncertainty in Biomedical Image Quantification2021.09MICCAI ChallengeProject
M&Ms-2M&Ms-2: Multi-Domain Cardiac MR Segmentation2021.09MICCAI ChallengeProject
Learn2Reg Abdomen MR-CTLearn2Reg 2021: Abdominal MR-CT Multi-Modal Registration2021.09MICCAI ChallengeProject
CrossMoDA2021CrossMoDA 2021: Unsupervised Domain Adaptation for Cross-Modality Vestibular Schwannoma Segmentation2021.09MICCAI ChallengeProject
WORDWORD: A Whole-Organ CT Dataset for Robust Multi-Organ Segmentation2021.10arXivPaperProject
MedMNIST v2MedMNIST v2 -- A large-scale lightweight benchmark for 2D and 3D biomedical image classification2021.10NaturePaperProject
CADACADA: Cerebral Aneurysm Detection and Analysis Challenge2022.04MICCAI ChallengeProject
CADA-ASCADA-AS: Aneurysm Segmentation Challenge2022.04MICCAI ChallengeProject
CADA-RRECADA-RRE: Rupture Risk Estimation for Cerebral Aneurysms2022.04MICCAI ChallengeProject
TotalSegmentatorTotalSegmentator: robust segmentation of 104 anatomic structures in CT images2022.06arXivPaperProject
AMOSAMOS: A Large-Scale Abdominal Multi-Organ Benchmark for Versatile Medical Image Segmentation2022.06NeurIPS 2022PaperProject
MSDThe medical segmentation decathlon2022.07Nature CommunicationsPaperProject
AutoPETThe AutoPET Challenge: Automated Lesion Segmentation in Whole-Body FDG-PET/CT2022.07MICCAI Challenge (autoPET2022)Project
UPENN-GBMUPENN-GBM: Multi-modal MRI Dataset for Glioblastoma Segmentation2022.07TCIA CollectionProject
KiPA22KiPA22: Kidney PArametric segmentation in contrast-enhanced CT2022.08MICCAI ChallengeProject
OLIVESOLIVES: A 3D OCT Dataset for Longitudinal Retinal Imaging2022.09arXivProject
PI-CAIPI-CAI: Prostate Imaging–Cancer AI Challenge2022.09MICCAI ChallengeProject
LAScarQS 2022LAScarQS 2022: Left Atrial Scar Quantification and Segmentation Challenge2022.09MICCAI ChallengeProject
CrossMoDA2022CrossMoDA 2022: Domain Adaptation for Vestibular Schwannoma Segmentation and Koos Grading2022.09MICCAI ChallengeProject
FeTA 2022FeTA 2022: Fetal Brain Tissue Segmentation at MICCAI2022.09MICCAI ChallengeProject
COSMOS 2022COSMOS 2022: Carotid Artery Vessel Wall Segmentation2022.09MICCAI ChallengeProject
cSeg-2022cSeg 2022: Cerebellum Segmentation Challenge2022.09MICCAI ChallengeProject
ISLES 2022ISLES 2022 – Acute and Subacute Ischemic Stroke Lesion Segmentation2022.09MICCAI ChallengeProject
InSTANCE2022InSTANCE 2022: Intracranial Hemorrhage Segmentation Challenge2022.09MICCAI ChallengePaperProject
Learn2Reg NLSTLearn2Reg 2022: Thoracic CT Registration with NLST2022.09MICCAI ChallengeProject
Shifts Challenge 2022Shifts 2022: Distribution Shifts in Multiple Sclerosis Lesion Segmentation2022.09MICCAI ChallengePaperProject
HECKTOR 22Overview of the HECKTOR challenge at MICCAI 2022: automatic head and neck tumor segmentation and outcome prediction in PET/CT2023MICCAI 2022PaperProject
LNDbLNDb challenge on automatic lung cancer patient management2023.03Medical Image AnalysisPaperProject
Semi-TeethSegSemi-TeethSeg: Semi-Supervised 3D Tooth Segmentation in CBCT/CT2023.04arXivProject
PARSE22Efficient automatic segmentation for multi-level pulmonary arteries: The parse challenge2023.04arXivPaperProject
STAGESTAGE: Longitudinal OCT Dataset for Glaucoma Progression2023.04DatasetProject
KiTS21The kits21 challenge: Automatic segmentation of kidneys, renal tumors, and renal cysts in corticomedullary-phase ct2023.07arXivPaperProject
AutoPET IIAutoPET-II: Ensemble-based Uncertainty-Aware Lesion Segmentation in Multi-Center FDG-PET/CT2023.07MICCAI Challenge (AutoPET-II)Project
MedMDTowards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical Data2023.08arXivPaperProject
ULS23ULS23 Challenge: Universal Lesion Segmentation in CT for Oncological Imaging2023.08MICCAI ChallengeProject
SegRap2023SegRap 2023: Nasopharyngeal Carcinoma Radiotherapy Segmentation Challenge2023.08MICCAI ChallengeProject
LNQ2023LNQ2023: Lymph Node Quantification in Chest CT2023.08MICCAI ChallengeProject
FLARE23FLARE 2023: A Federated Learning Challenge for Abdominal Multi-Organ Segmentation2023.09MICCAI ChallengeProject
CrossMoDA2023CrossMoDA 2023: Multi-Center Domain Adaptation for VS Segmentation2023.09MICCAI ChallengeProject
ATLAS2023ATLAS 2023: Liver Tumor Segmentation Challenge2023.09MICCAI ChallengePaperProject
SMILE-UHURA2023SMILE-UHURA 2023: Small Vessel Disease Lesion Segmentation2023.09MICCAI ChallengePaperProject
CAS2023CAS 2023: Brain Structure Segmentation Benchmark2023.09MICCAI ChallengeProject
CROWN2023CROWN 2023: White Matter Hyperintensity and Other Pathology Classification2023.09MICCAI ChallengeProject
SLCNSLCN: Structural Lesion and Connectivity in Neurodevelopmental Disorders2023.09MICCAI ChallengeProject
ToothFairy2023ToothFairy: 3D CBCT Dataset for Inferior Alveolar Nerve Segmentation2023.09MICCAI ChallengeProject
XPRESS2023XPRESS 2023: X-ray Phase-Contrast CT Neuroanatomy Segmentation2023.09MICCAI ChallengePaperProject
Learn2Reg ThoraxCBCTLearn2Reg 2023: Thorax CBCT/FBCT Deformable Registration2023.09MICCAI ChallengeProject
TDSC-ABUS2023TDSC-ABUS2023: Automated Breast Ultrasound Segmentation Challenge2023.09MICCAI ChallengePaperProject
MVSeg-3DTEE2023MVSeg-3DTEE2023: Mitral Valve Segmentation from 3D TEE2023.09MICCAI ChallengeProject
RegPro2023RegPro 2023: Prostate MR-US Registration Challenge2023.09MICCAI ChallengeProject
KiTS23KiTS23: Kidney and Kidney Tumor Segmentation with Comprehensive Clinical Annotations2023.10arXivProject
WBMR-NFWBMR-NF: Whole-Body MRI for Neurofibromatosis2023.11DatasetProject
GAMMAGAMMA Challenge: Glaucoma Assessment with Multi-Modality Data2023.12MICCAI ChallengePaperProject
ATM'22ATM'22: Airway Tree Modeling in Thoracic CT2023.12MICCAI ChallengePaperProject
INSPECTINSPECT: A Multimodal Dataset for Pulmonary Embolism Diagnosis and Prognosis2023.12NeurIPS 2023PaperProject
HaN-SegHaN-Seg 2023: Head and Neck Organ at Risk Segmentation Challenge2024MICCAI ChallengePaperProject
VALDOWhere is VALDO? Vascular Lesions Detection and Segmentation Challenge2024.01MICCAI ChallengePaperProject
BIMCV-RBIMCV-R: Large-Scale Thoracic CT Reconstruction Benchmark2024.01MICCAI24PaperProject
IXIInformation eXtraction from Images (IXI) Dataset2024.01DatasetProject
RAOSRAOS: A Large-Scale Radiotherapy Abdominal Organ Segmentation Dataset2024.01MICCAI24PaperProject
ISLES 2024Ischemic Stroke Lesion Segmentation Challenge 2024 (ISLES 2024)2024.02MICCAI ChallengeProject
AbdomenAtlasAbdomenAtlas-20K: A Large-Scale Benchmark for Abdominal Multi-Organ Segmentation in CT2024.02arXivPaperProject
OpenMindOpenMind: Large-Scale Head-and-Neck MR Dataset for Foundation Models2024.02arXivProject
TriALS2024TriALS 2024: Liver Tumor Segmentation and Outcome Prediction – Task 12024.03MICCAI24Project
LAScarQS++ 2024LAScarQS++ 2024: Multi-Center Atrial Scar Segmentation2024.03CARE Workshop (MICCAI)Project
MyoPSMyoPS: A Benchmark of Myocardial Pathology Segmentation Combining Three-Sequence Cardiac Magnetic Resonance Images# MyoPS: A Benchmark of Myocardial Pathology Segmentation Combining Three-Sequence Cardiac Magnetic Resonance Images2024.03CARE WorkshopPaperProject
WHS++ 2024WHS++ 2024: Multi-Center Whole Heart Segmentation2024.03CARE WorkshopProject
AMOS-MMAMOS-MM: Multi-Phase Abdominal CT Benchmark for Translation and Synthesis2024.03arXivPaperProject
CT2RepCT2Rep: Automated Radiology Report Generation for 3D Medical Imaging2024.03MICCAI 2024PaperProject
TotalSegmentator MRITotalSegmentator MRI: Whole-body MRI Segmentation of 150 Structures2024.04arXivPaperProject
RadGenome-ChestCTRadGenome-Chest CT: a grounded vision-language dataset for chest CT analysis2024.04arXivPaperProject
M3DM3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models2024.04ICLR 2025PaperProject
AIIB23AIIB23: Airway Inflammation Imaging Biomarkers Challenge2024.06MICCAI ChallengePaperProject
CT-3DRRGArgus: Benchmarking and Enhancing Vision-Language Models for 3D Radiology Report Generation2024.06arXivPaper—
RadGenome-Brain MRIAutoRG-Brain: Grounded Report Generation for Brain MRI2024.10MICCAI 2024PaperProject
MedShapeNetMedShapeNet -- A Large-Scale Dataset of 3D Medical Shapes for Computer Vision2024.12Biomedizinische TechnikPaperProject
RadA-BenchPlatHow well can modern LLMs act as agent cores in radiology environments?2024.12arXivPaper
MedVL-CT69KLarge-scale and Fine-grained Vision-language Pre-training for Enhanced CT Image Understanding2025.01ICLR 2025PaperProject
TriadTriad: Vision Foundation Model for 3D Magnetic Resonance Imaging2025.02arXivPaper
3D-BrainCTTowards a holistic framework for multimodal LLM in 3D brain CT radiology report generation2025.03Nat. Commun.Paper—
PENGWIN2024-Task1PENGWIN 2024: Pelvic Fracture Segmentation in Trauma CT2025.04MICCAI ChallengePaperProject
RibFracDeep rib fracture instance segmentation and classification from ct on the ribfrac challenge2025.04IEEEPaperProject
DeepTumorVQAAre Vision Language Models Ready for Clinical Diagnosis? A 3D Medical Benchmark for Tumor-centric Visual Question Answering2025.05NeurIPS 2025PaperProject
NOVANOVA: A Benchmark for Anomaly Localization and Clinical Reasoning in Brain MRI2025.05NeurIPS 2025Paper
LingshuLingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning2025.06arXivPaperProject
ReXGroundingCTReXGroundingCT: A 3D Chest CT Dataset for Segmentation of Findings from Free-Text Reports2025.07arXivPaperProject
ViPET-ReportGenToward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation2025.12NeurIPS 2025PaperProject
3D-RAD3D-RAD: A Comprehensive 3D Radiology Med-VQA Dataset with Multi-Temporal Analysis and Diverse Diagnostic Tasks2025.12NeurIPS 2025PaperProject
MR-RATEMR-RATE: A Vision-Language Foundation Model and Dataset for Magnetic Resonance Imaging2026——Project
CT-RATEGeneralist Foundation Models from a Multimodal Dataset for 3D Computed Tomography2026.02NaturePaperProject
CT-FlowBenchCT-FlowBench: Benchmark for CT interpretation workflow and tool-use2026.03arXivPaper
MerlinMerlin: A Vision Language Foundation Model for 3D Computed Tomography2026.03NaturePaperProject
Gastric-XGastric-X: A Multimodal Multi-Phase Benchmark Dataset for Advancing Vision-Language Models in Gastric Cancer Analysis2026.03CVIPPR 2026Paper—
SpatialMedBeyond Medical Diagnostics: How Medical Multimodal Large Language Models Think in Space2026.03arXivPaper
BONBID-HIE2023BONBID-HIE 2023: Neonatal Hypoxic-Ischemic Encephalopathy Lesion Segmentation2026.04MICCAI ChallengePaperProject
SGMRI-VQABeyond a Single Frame: Multi-Frame Spatially Grounded Reasoning Across Volumetric MRI2026.04arXivPaper
Curia-2Curia-2: Scaling Self-Supervised Learning for Radiology Foundation Models2026.04arXivPaper
CT-SpatialVQALost in Volume: The CT-SpatialVQA Benchmark for Evaluating Semantic-Spatial Understanding of 3D Medical Vision-Language Models2026.05arXivPaper
Med-StepBenchMed-StepBench: A Hierarchical Reasoning Framework for Evaluating Hallucinations in Medical Vision-Language Models2026.05arXivPaper
DeepTumorVQA-HDeepTumorVQA: A Hierarchical 3D CT Benchmark for Stage-Wise Evaluation of Medical VLMs and Tool-Augmented Agents2026.05arXivPaper
ABRAABRA: Agent Benchmark for Radiology Applications2026.05arXivPaper
RadSaFE-200Safety and Accuracy Follow Different Scaling Laws in Clinical Large Language Models2026.05arXivPaper
Oncology VQA BenchmarkAutomated Report-Derived Oncology VQA Benchmark for Evaluating Vision-Language Models on 3D Medical Imaging2026.06arXivPaper
Abdomen-NCCT BenchmarkA Multi-Center Benchmark for Abdominal Disease Diagnosis and Report Generation from Non-Contrast CT2026.06arXivPaper
RadOT-EvalRadOT-Eval: Auditable Structured-Evidence Transport for Radiology Report Evaluation2026.06arXivPaper
ReportQAReportQA: QA-Based Radiology Report Evaluation2026.06arXivPaper
CORTEXCORTEX: A Structured Reasoning Benchmark for Trustworthy 3D Chest CT MLLMs2026.06arXivPaper
MedCTAMedCTA: A Benchmark for Clinical Tool Agents2026.06arXivPaper
Lung CT FM BenchmarkFoundation Models vs. Radiomics for Lung Computed Tomography: A Benchmark of Feature Extractors, Classification Heads, and Segmentation Choices2026.07arXivPaperProject
Brain Oncology 3D MRI-TextMulti-LLM Collaborative MRI Report Generation for Visual Instruction Tuning in Brain Oncology2026.07arXivPaper-
COBRA2026COBRA2026: a large-scale multicenter pelvic cone-beam computed tomography projection dataset2026.07arXivPaperProject
GLI-ALGLI-AL: A Multi-Modal Glioma MRI Label Resource with Unified Anatomy-Lesion Labels2026.07arXivPaperProject

Contributors

Saint-lsy

16 commits

yezanting

13 commits

Leo-lxin

10 commits

Saint-lsy/Awesome-Volumetric-Radiology-AI

Volumetric Radiology AI in the Era of Multimodal Large Language Models

23

39 commits

updated Aug 24, 2026

See the code

README

Volumetric Radiology AI in the Era of Multimodal Large Language Models

arXiv:2608.20549 Download PDF

The official resource repository for our review of volumetric radiology foundation models, multimodal large language models, agentic systems, and evaluation benchmarks.

Timeline showing the co-development of volumetric radiology foundation models and agentic systems

Co-development of volumetric radiology foundation models and agentic systems.

Contents


Foundation Models

Three-stage landscape of volumetric radiology foundation models

Foundation-model landscape: volumetric self-supervised pre-training, vision-language alignment, and MLLM-based interpretation.

Self-Supervised Pre-training

MethodTitleVenueDatePaperProject
Models GenesisMedical Image Analysis21.02PaperProject
UniMiSSUniMiSS: Universal Medical Self-Supervised Learning via Breaking Dimensionality BarrierECCV2022&TPAMI21.12PaperProject
Swin-UNETRSelf-Supervised Pre-Training of Swin Transformers for 3D Medical Image AnalysisCVPR 2022CVPR 2022PaperProject
PCRLv2A Unified Visual Information Preservation Framework for Self-supervised Pre-training in Medical Image AnalysisIEEE TPAMI2023.01PaperProject
MIM-Med3DMasked Image Modeling Advances 3D Medical Image AnalysisWACV 20232023.01PaperProject
GVSLGeometric Visual Similarity Learning in 3D Medical Image Self-supervised Pre-trainingCVPR 20232023.03PaperProject
M3AEM3AE: Multimodal Representation Learning for Brain Tumor Segmentation with Missing ModalitiesAAAI 20232023.03PaperProject
HybridMIMHybridMIM: A Hybrid Masked Image Modeling Framework for 3D Medical Image SegmentationIEEE JBHI2024.04PaperProject
VoCoLarge-Scale 3D Medical Image Pre-Training With Geometric Context PriorsCVPR 2024 / IEEE TPAMI2024.06PaperProject
CDSSL-P3DCross-Dimensional Medical Self-Supervised Representation Learning Based on a Pseudo-3D TransformationMICCAI 242024.10Paper-
MDMMasked Deformation Modeling for Volumetric Brain MRI Self-Supervised Pre-TrainingIEEE TMI2024.12PaperProject
DAEDisruptive Autoencoders: Leveraging Low-level features for 3D Medical Image Pre-trainingMIDL 20242024.06PaperProject
Unified 3D MRI Representations via Sequence-Invariant Contrastive LearningarXiv2025.01PaperProject
GzPTImproving Self-Supervised Medical Image Pre-Training by Early Alignment With Human Eye Gaze InformationAAAI24/IEEE TMI2025.01PaperProject
FM-HCT3D Foundation AI Model for Generalizable Disease Detection in Head Computed TomographyarXiv2025.02PaperProject
MedVAE: Efficient Automated Interpretation of Medical Images with Large-Scale Generalizable AutoencodersMIDL 20252025.02PaperProject
MiMMiM: Mask in Mask Self-Supervised Pre-Training for 3D Medical Image AnalysisIEEE TMI2025.04Paper-
BrainMVPBrainMVP: Multi-modal Vision Pre-training for Brain MRI AnalysisCVPR 2025 (Highlight)2025.06PaperProject
SPECTREScaling Self-Supervised and Cross-Modal Pretraining for Volumetric CT TransformersCVPR 20262025.11PaperProject
SubForeHU-based Foreground Masking for 3D Medical Masked Image ModelingMICCAI252025.10PaperProject
3DINOA generalizable 3D framework and model for self-supervised learning in medical imagingnpj Digital Medicine2025.11PaperProject
TotalFMTotalFM: An Organ-Separated Framework for 3D-CT Vision Foundation Modelsarxiv2026.01Paper-
Curia-2Curia-2: Scaling Self-Supervised Learning for Radiology Foundation Modelsarxiv2026.02Paper-
OCTCube-MA 3D multimodal optical coherence tomography foundation model for retinal and systemic diseases with cross-cohort and cross-device validationNature Biomedical Engineering2026.04PaperProject
NeuroSTORMTowards a general-purpose foundation model for fMRI analysisNBME2026.03PaperProject
TriadVision foundation model for 3D magnetic resonance imaging segmentation, classification, and registrationMedIA2026.05Paper-
Foundation-VAEFoundation VAEs for 3D CT Reconstruction, Augmentation, and GenerationICML 20262026.05PaperProject
CoralBayCoralBay: A Self-Supervised CT Foundation ModelarXiv2026.06Paper-
How Much MRI Preprocessing Is Enough? A Cost-Utility Study for Brain MRI Foundation ModelsarXiv2026.06Paper-
NeuroVFMHealth system learning enables generalist neuroimaging modelsNature Medicine2026.07PaperProject
BrainFIBREBrainFIBRE: A Foundation Model via Information Decomposition for Brain MicrostructureECCV 2026 / arXiv2026.07Paper-
BoneCoT / BoneFMBoneCoT: multicentre validation of a whole-body skeleton foundation model for bone metastases guided by clinician-derived chain of thoughtNature Biomedical Engineering2026.07PaperProject
Cardiac CT FMA Unified Framework for Comprehensive Cardiac CT Segmentation and Phenotyping: Human-in-the-Loop Data Annotation, Vision Foundation Model Development, Multicenter Evaluation and Clinical ValidationarXiv2026.07Paper-
COJEPAContrastive Joint-Embedding Prediction for Representation Learning in Structural MRIarXiv2026.07Paper-
BrainNextBrainNext: A General-Purpose Self-Supervised Foundation Model for Brain MRI AnalysisarXiv2026.07Paper-
OrganLensOrganLens: Organ-Specific Representation Learning for CT Foundation ModelsarXiv2026.07Paper-
Rad-JEPA 3DRad-JEPA 3D: Radiology Joint-Embedding Predictive Model for 3D Computed TomographyarXiv2026.07Paper-

Contrastive Learning (CLIP-based Methods)

MethodDimensionTitleVenueDatePaperProject
MedCLIP2DMedCLIP: Contrastive Learning from Unpaired Medical Images and TextEMNLP 20222022.12PaperProject
PubMedCLIP2DPubMedCLIP: How Much Does CLIP Benefit Visual Question Answering in the Medical Domain?EACL 2023 (Findings)2023.05PaperProject
MedBLIP2DMedBLIP: Bootstrapping Language-Image Pre-training from 3D Medical Images and Textsarxiv2023.05Paper-
CLIP-Lung2DCLIP-Lung: Textual Knowledge-Guided Lung Nodule Malignancy PredictionMICCAI 20232023.10PaperProject
BioMedCLIP2DBiomedCLIP: A Multimodal Biomedical Foundation Model Pretrained from Fifteen Million Scientific Image-Text PairsNEJM AI 20242024.01PaperProject
PMC-CLIP2DPMC-CLIP: Contrastive Language-Image Pre-training using Biomedical DocumentsMICCAI 20232023.10PaperProject
UniMedCLIP2DUniMed-CLIP: Towards a Unified Image-Text Pretraining Paradigm for Diverse Medical Imaging ModalitiesArxiv2024.12PaperProject
ConceptCLIP2DConceptCLIP: Towards Trustworthy Medical AI via Concept-Enhanced Contrastive Language-Image Pre-trainingarXiv2025.01PaperProject
MMKD-CLIP2DUnifying Biomedical Vision-Language Expertise: Towards a Generalist Foundation Model via Multi-CLIP Knowledge DistillationArxiv2025.06Paper-
RadiSimCLIP2DRadiSimCLIP: A Radiology Vision-Language Model Pretrained on Simulated Radiologist Learning Dataset for Zero-Shot Medical Image UnderstandingMICCAI 2025 Workshop2025.10PaperProject
UniBrain3DUniBrain: Universal Brain MRI Diagnosis with Hierarchical Knowledge-enhanced Pre-trainingComputerized Medical Imaging and Graphics2023.09PaperProject
CT2Rep3DCT2Rep: Automated Radiology Report Generation for 3D Medical ImagingMICCAI 20242024.03PaperProject
CT-CLIP3DGeneralist Foundation Models from a Multimodal Dataset for 3D Computed TomographyNat Biomed Eng2024.03PaperProject
CT-GLIP3DCT-GLIP: 3D Grounded Language-Image Pretraining with CT Scans and Radiology Reports for Full-Body ScenariosarXiv2024.04Paper-
RadCLIP3DRadCLIP: Enhancing Radiologic Image Analysis Through Contrastive Language–Image PretrainingTNNLS2024.03Paper-
Percival3DA Pan-Organ Vision-Language Model for Generalizable 3D CT RepresentationsmedRxiv2025.07Paper-
OpenVocabCT3DTowards Universal Text-driven CT Image SegmentationarXiv2025.03Paper-
fVLM3DLarge-scale and Fine-grained Vision-language Pre-training for Enhanced CT Image UnderstandingICLR 20252025.03PaperProject
HLIP3DTowards Scalable Language-Image Pre-training for 3D Medical Imagingarxiv2025.05PaperProject
RadZero3D3DRadZero3D: Bridging Self-Supervised Video Models and Medical Vision-Language Alignment for Zero-Shot Chest CT InterpretationICCV 2025 Workshop2025.10Paper-
T3D3DT3D: Advancing 3D Medical Vision-Language Pre-training by Learning Multi-View Visual ConsistencyICCV 2025 Workshop2025.10Paper-
ViSD-Boost3DBoosting Vision Semantic Density with Anatomy Normality Modeling for Medical Vision-language Pre-trainingICCV 20252025.08PaperProject
VELVET-Med3DVELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in MedicinearXiv2025.08Paper-
MedVista3D3DMedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and ReportingarXiv2025.09Paper-
COLIPRI3DComprehensive Language-Image Pre-training for 3D Medical Image UnderstandingarXiv2025.10PaperProject
MPS-CT3DMore performant and scalable: Rethinking contrastive vision-language pre-training of radiology in the LLM eraMICCAI 20252025.10PaperProject
PET-CLIP Captioner3DLocation-Guided Automated Lesion Captioning in Whole-body PET/CT ImagesMICCAI 20252025.10Paper-
MR-CLIP3DMetadata-Aligned 3D MRI Representations for Contrast Understanding and Quality ControlarXiv2025.11Paper-
BrgSA3DBridged Semantic Alignment for Zero-shot 3D Medical Image DiagnosisarXiv2025.11PaperProject
SCALE-VLP3DSCALE-VLP: Soft-Weighted Contrastive Volumetric Vision-Language Pre-training with Spatial-Knowledge SemanticsarXiv2025.11Paper-
SPECTRE3DScaling Self-Supervised and Cross-Modal Pretraining for Volumetric CT TransformersCVPR 20262025.11PaperProject
Pillar-03DPillar-0: A New Frontier for Radiology Foundation Modelsarxiv2025.11PaperProject
BTB3D3DBetter Tokens for Better 3D: Advancing Vision-Language Modeling in 3D Medical ImagingNeurIPS 20252025.12PaperProject
NeuroVFM3DNeuroVFM: A Contrastive Vision-Language Model for Medical Reasoning in Alzheimer's Disease DiagnosisWACV 2026 Workshops2026.01Paper-
TotalFM3DTotalFM: An Organ-Separated Framework for 3D-CT Vision Foundation Modelsarxiv2026.01Paper-
MG-3D3DMG-3D: Multi-Grained Knowledge-Enhanced 3D Medical Vision-Language Pre-trainingMedical Image Analysis (MedIA)2026.01PaperProject
MedMAP3D3D Modality-Aware Pre-training for Vision-Language Model in MRI Multi-organ Abnormality Detectionarxiv2026.02PaperProject
Prima3DLearning neuroimaging models from health system-scale dataNature Biomedical Engineering2026.02PaperProject
SigVLP3DSigVLP: Sigmoid Volume-Language Pre-Training for Self-Supervised CT-Volume Adaptive Representation LearningarXiv2026.02Paper-
RadFinder3DLearning to Read Where to Look: Disease-Aware Vision–Language Pretraining for 3D CTarXiv2026.03PaperProject
Decipher-MR3DDecipher-MR: A Vision-Language Foundation Model for 3D MRI Representationsnpj Digital Medicine2026.04Paper-
3DCLIP Architecture for Abdominal CT Image–Text Alignment and Zero-Shot Learning: Investigating Batch Composition and Data Scalingarxiv2026.04Paper-
ASAP3DASAP: Advancing Medical Volumetric Representation Learning with Anatomy-aware Semantically-adaptive Pre-trainingarXiv2026.05Paper-
GLeVE3DGLeVE: Graph-Guided Lesion Grounding with Proposal Verification in 3D CTarXiv2026.05Paper-
CA-GCL3DCA-GCL: Cross-Anatomy Global-Local Contrastive Learning for Robust 3D Medical Image UnderstandingarXiv2026.05Paper-
SegReg-Rep3DSegReg-Rep: Region-Aware Vision-Language Alignment for Fine-Grained Radiology Report Generation from 3D Medical ImagesIEEE TPRMS2026.05PaperProject
GLINT3DGLINT: Sparsely Gated Vision-Language Alignment for Fine-Grained Radiology RepresentationsarXiv2026.06Paper-
RadGrounder2DScalable Training of Spatially Grounded 2D Vision-Language Models for RadiologyarXiv2026.06Paper-
RenalCLIP3DA Disease-Centric Vision-Language Foundation Model for Precision Oncology in Kidney CancerNature Communications2026.06Paper-
Jolia / ConQuer3DJolia: Concept-Level Vision-Language Alignment for 3D CT Contrastive LearningarXiv2026.06Paper-
3DDisease-Centric Vision-Language Pretraining with Hybrid Visual Encoding for 3D Computed TomographyarXiv2026.06Paper-
MedReCo3DA Vision-language Framework for Comparative Reasoning in RadiologyarXiv2026.06Paper-
SuG3DSuper-Generalist: Towards Comprehensive and Accurate Medical Image Understanding via Generalist-Specialist SynergyarXiv2026.07Paper-
OKA-CT3DLearning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report KnowledgearXiv2026.07Paper-
OCP-CT3DFine-Grained Vision-Language Pretraining with Organ-Conditioned Pattern Tokens for CT UnderstandingarXiv2026.07Paper-
MseaCL3DMultimodal Semantic-Aware Contrastive Learning For False Negative Mitigation in 3D Medical ImagingarXiv2026.07Paper-
CARVE3DWhen Can Test-Time Adaptation Help Zero-Shot CT Vision-Language Models?arXiv2026.07Paper-
ACA3DAnatomy Contextualized Adaption of CT Foundation ModelsarXiv2026.07Paper-
Spectrum3DLearning How Much, Not Just What: Cross-Patient Burden Order for CT Vision-Language PretrainingarXiv2026.07Paper-
SCOPE3DSemantically Calibrated Evidence Composition for CT Vision-Language LearningarXiv2026.07Paper-

MLLM-based Methods

2D-only

MethodTitleVenueDatePaperProject
BiomedGPTBiomedGPT: A Generalist Vision-Language Foundation Model for Diverse Biomedical TasksNat Med 20242023.05PaperProject
MedVInTPMC-VQA: Visual Instruction Tuning for Medical Visual Question AnsweringArxiv/ Communications Medicine2023.05/2024.12Paper-
LLaVA-MedLLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One DayNeurIPS 20232023.06PaperProject
Med-FlamingoMed-Flamingo: a Multimodal Medical Few-shot LearnerML4H 20232023.07PaperProject
Med-PaLM MTowards Generalist Biomedical AINEJM AI 2024 / arXiv2023.07PaperProject
Qilin-Med-VLQilin-Med-VL: Towards Chinese Large Vision-Language Model for General HealthcarearXiv2023.10Paper-
R2GenGPTR2GenGPT: Radiology Report Generation with Frozen LLMsMeta-Radiology2023.11Paper-
BiRDA Refer-and-Ground Multimodal Large Language Model for BiomedicineMICCAI 20242024.06Paper-
HuatuoGPT-VisionTowards Injecting Medical Visual Knowledge into Multimodal LLMs at ScaleEMNLP242024.06PaperProject
Llama3-MedAdvancing High Resolution Vision-Language Models in BiomedicinearXiv2024.06PaperProject
MiniGPT-MedMiniGPT-Med: Large Language Model as a General Interface for Radiology DiagnosisarXiv2024.07PaperProject
TinyLLaVA-MedDemocratizing MLLMs in Healthcare: TinyLLaVA-Med for Efficient Healthcare Diagnostics in Resource-Constrained SettingsMICCAI242024.09Paper-
Med-MoEMed-MoE: Mixture of Domain-Specific Experts for Lightweight Medical Vision-Language ModelsEMNLP242024.09PaperProject
GMAI-VLGMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AIarXiv/AAAI262024.11PaperProject
BiMediX2BiMediX2: Bio-Medical EXpert LMM for Diverse Medical ModalitiesFindings of EMNLP 20252025.01PaperProject
HealthGPTHealthGPT: A Medical Large Vision-Language Model for Unifying Comprehension and Generation via Heterogeneous Knowledge AdaptationICML 20252025.02PaperProject
MedVLM-R1MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement LearningMICCAI252025.02PaperProject
Med-R1Med-R1: Reinforcement Learning for Generalizable Medical Reasoning in Vision-Language ModelsTMI2025.03PaperProject
OmniV-MedOmniV-Med: Scaling Medical Vision-Language Model for Universal Visual UnderstandingarXiv2025.04Paper-
UniBioMedUniBiomed: A Universal Foundation Model for Grounded Biomedical Image InterpretationarXiv2025.04PaperProject
MedRegAMedRegA: Interpretable Bilingual Multimodal Large Language Model for Diverse Biomedical TasksICLR252025.04PaperProject
UMed-LVLMImproving Medical Large Vision-Language Models with Abnormal-Aware FeedbackACL 20252025.05Paper-
QoQ-MedQoQ-Med: Building Multimodal Clinical Foundation Models with Domain-Aware GRPO TrainingarXiv2025.06PaperProject
LingshuLingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and ReasoningarXiv2025.06PaperProject
MedGemmaMedGemma Technical ReportarXiv2025.07PaperProject
Critus-VCitrus-V: Advancing Medical Foundation Models with Unified Medical Image Grounding for Clinical ReasoningarXiv2025.09PaperProject
MedPLIBTowards a Multimodal Large Language Model with Pixel-Level Insight for BiomedicineAAAI252025.10PaperProject
OctoMedOctoMed: Data Recipes for State-of-the-Art Multimodal Medical ReasoningarXiv2025.11PaperProject
MedMOMedMO: Grounding and Understanding Multimodal Large Language Model for Medical ImagesarXiv2026.02PaperProject
MediX-R1MediX-R1: Open Ended Medical Reinforcement LearningarXiv2026.02PaperProject
MEDIC-ADMEDIC-AD: Towards Medical Vision-Language Model’s Clinical IntelligenceCVPR26 Oral2026.03PaperProject
MedVRMedVR: Annotation-Free Medical Visual Reasoning via Agentic Reinforcement LearningICLR 262026.04PaperProject

Unified 2D + Serialized-3D

MethodTitleVenueDatePaperProject
RadFMTowards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical DataNat Commun 2025 / arXiv2023.08PaperProject
Med-GeminiAdvancing Multimodal Medical Capabilities of GeminiNat Med 2025 / arXiv2024.05PaperProject
Med-2E3Med-2E3: A 2D-Enhanced 3D Medical Multimodal Large Language ModelBIBM252024.11Paper-
Hulu-MedHulu-Med: A Transparent Generalist Model towards Holistic Medical Vision-Language UnderstandingarXiv2025.10PaperProject
Fleming-VLFleming-VL: Towards Universal Medical Visual Reasoning with Multimodal LLMsarXiv2025.11PaperProject
MedM-VLMedM-VL: What Makes a Good Medical LVLM?International Workshop on Agentic AI for Medicine 20252025.09PaperProject
CTInstructCTInstruct: Towards Unified 3D CT Understanding via Instruction TuningAAAI 262026.01PaperProject
A data-efficient 3D medical vision-language model using only a 2D encoderScientific report2026.02Paper-
MedPrunerMedPruner: Training-Free Hierarchical Token Pruning for Efficient 3D Medical Image Understanding in Vision-Language ModelsMICCAI20262026.03Paper-
PhotonPhoton: Speedup Volume Understanding with Efficient Multimodal Large Language ModelsICLR 262026.03PaperProject
OmniCTOmniCT: Towards a Unified Slice-Volume LVLM for Comprehensive CT AnalysisICLR 262026.03PaperProject
MedGemma1.5MedGemma 1.5 Technical ReportarXiv2026.04PaperProject
TGH-MoEAdapting 2D Multi-Modal Large Language Model for 3D CT Image AnalysisarXiv2026.04Paper-
Brain-AdapterBrain-Adapter: A Dual-Stream Vision-Language MIL Framework for Comprehensive 3D CT Diagnosis of Acute Intracranial PathologiesarXiv2026.06Paper-
UniReason-MedUniReason-Med: A Shared Grounded Reasoning Interface for 2D-to-3D Transfer in Medical VQAarXiv2026.06PaperProject
MedReCo-VLMA Vision-language Framework for Comparative Reasoning in RadiologyarXiv2026.06Paper-
RadSightRadSight: Towards Perceptually Reliable Multimodal Radiology Image UnderstandingarXiv2026.07PaperProject
ClinFusionClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical UnderstandingarXiv2026.07PaperProject
HounsfieldTowards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language ModelsarXiv2026.07PaperProject
MedARCMedARC: Training-Free Adaptive Redundancy Compression of Visual Tokens for 3D Medical Vision-Language ModelsarXiv2026.07Paper-
ORCAORCA: ORgan-Centroid Aggregation for Training-Free 3D CT Visual Token CompressionarXiv2026.07PaperProject

Native 3D-Volume

MethodTitleVenueDatePaperProject
CT2RepCT2Rep: Automated Radiology Report Generation for 3D Medical ImagingMICCAI 2024 / arXiv2024.03PaperProject
Dia-LLaMADia-LLaMA: Towards Large Language Model-driven CT Report GenerationMICCAI2024.03Paper-
M3D-LaMedM3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language ModelsICLR 2025 / arXiv2024.04PaperProject
MerlinMerlin: A Computed Tomography Vision-Language Foundation Model and DatasetNature 2026 / arXiv2024.06PaperProject
BrainGPTTowards a Holistic Framework for Multimodal Large Language Models in Three-dimensional Brain CT Report GenerationNat Commun 2025 / arXiv2024.07PaperProject
3D-CT-GPT3D-CT-GPT: Generating 3D Radiology Reports through Integration of Large Vision-Language ModelsarXiv2024.09Paper-
E3D-GPTE3D-GPT: Enhanced 3D Visual Foundation for Medical Vision-Language ModelarXiv2024.10Paper-
Reg2RGLarge Language Model with Region-guided Referring and Grounding for CT Report GenerationarXiv2024.11Paper-
MS-VLMRead Like a Radiologist: Efficient Vision-Language Model for 3D Medical Imaging InterpretationarXiv2024.12Paper-
MEPNetMEPNet: Medical Entity-balanced Prompting Network for Brain CT Report GenerationarXiv2025.03Paper-
Med3DVLMMed3DVLM: An Efficient Vision-Language Model for 3D Medical Image AnalysisIEEE JBHI 2025 / arXiv2025.03PaperProject
HSENetHSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language UnderstandingarXiv2025.06Paper-
MedRegion-CTMedRegion-CT: Region-Focused Multimodal LLM for Comprehensive 3D CT Report GenerationarXiv2025.06Paper-
mpLLMMultimodal LLM With Hierarchical Mixture-of-Experts for VQA on 3D Brain MRIarXiv2025.09Paper-
3DReasonKnee3DReasonKnee: Advancing Grounded Reasoning in Medical Vision Language ModelsarXiv2025.10Paper-
PETARPETAR: Localized Findings Generation with Mask-Aware Vision-Language Modeling for PET Automated ReportingarXiv2025.10Paper-
BTB3DBetter Tokens for Better 3D: Advancing Vision-Language Modeling in 3D Medical ImagingNeurlPS 20252025.10PaperProject
PETRG-3DVision-Language Models for Automated 3D PET/CT Report GenerationarXiv2025.11Paper-
CTest-MetricCTest-Metric: A Unified Framework to Assess Clinical Validity of Metrics for CT Report GenerationISBI20262026.01Paper-
Brain3DBrain3D: Brain Report Automation via Inflated Vision Transformers in 3DarXiv2026.02PaperProject
Med3D-R1Med3D-R1: Incentivizing Clinical Reasoning in 3D Medical Vision-Language Models for Abnormality DiagnosisarXiv2026.02Paper-
LoV3DLoV3D: Grounding Cognitive Prognosis Reasoning in Longitudinal 3D Brain MRI via Regional Volume AssessmentsarXiv2026.03Paper-
Ker-VLJEPA-3BCurriculum-Driven 3D CT Report Generation via Language-Free Visual Grafting and Zone-Constrained CompressionarXiv2026.03PaperProject
U-VLMU-VLM: Hierarchical Vision Language Modeling for Report GenerationarXiv2026.02PaperProject
CT-CHATGeneralist foundation models from a multimodal dataset for 3D computed tomographyNature Biomedical Engineering2026.02PaperProject
MedVL-SAM2MedVL-SAM2: A unified 3D medical vision–language model for multimodal reasoning and prompt-driven segmentationarXiv2026.01Paper-
BoiDEnhancing 3D medical multi-modal large language models with integrated human body priors for computed tomographyPattern Recognition2026.04Paper-
DCP-PDEnhancing Fine-Grained Spatial Grounding in 3D CT Report Generation via Discriminative Guidancearxiv2026.04Paper-
SegReg-RepSegReg-Rep: Region-Aware Vision-Language Alignment for Fine-Grained Radiology Report Generation from 3D Medical ImagesIEEE TPRMS2026.05PaperProject
CLarGenGenerating Reports or Repeating Templates? Measuring and Mitigating Template Collapse in 3D CT Report GenerationarXiv2026.05Paper-
TIF-GRPORegulating Anatomy-Aware Rewards via Trajectory-Integral Feedback for Volumetric Computed Tomography AnalysisarXiv2026.05Paper-
RAD3D-PrefixRevisiting LLM Adaptation for 3D CT Report Generation: A Study of Scaling and Diagnostic PriorsarXiv2026.06Paper-
E-MRLE-MRL: Cross-view Aligned Evidence-driven Multimodal Reinforcement Learning for Reliable 3D Tumor AnalysisarXiv2026.06Paper-
MRI2RepMRI2Rep: Autoregressive Structured Report Generation for 3D Liver MRIarXiv2026.06Paper-
NeuroVFMHealth system learning enables generalist neuroimaging modelsNature Medicine2026.07PaperProject
PIPAPIPA: Prior-Driven Prompting with Diagnosis-Oriented Retrieval-Augmentation for 3D Radiology Report GenerationIEEE TMI2026.07PaperProject
MonteRETMonteRET: AI Agent Enhancing Multimodal LLMs with Multi-granularity Knowledge Retrieval for Chest CT Report GenerationarXiv2026.07Paper-
Multi-LLM MRIMulti-LLM Collaborative MRI Report Generation for Visual Instruction Tuning in Brain OncologyarXiv2026.07Paper-

Agentic Systems

Core modules of an agentic workflow for 3D radiology analysis

Core modules of agentic 3D radiology workflows: reasoning and planning, tool-augmented perception, memory and context, and workflow collaboration.

MethodTitleVenueDatePaperProject
MDAgentsMDAgents: An Adaptive Collaboration of LLMs for Medical Decision-MakingNeurIPS 20242024.12 (arXiv: 2024.04)PaperProject
MMedAgentMMedAgent: Learning to Use Medical Tools with Multi-modal AgentFindings of EMNLP 20242024.11 (arXiv: 2024.07)PaperProject
MedAgent-ProMedAgent-Pro: Towards Evidence-based Multi-modal Medical Diagnosis via Reasoning Agentic WorkflowICLR 20262025.03PaperProject
DOLAAutonomous Radiotherapy Treatment Planning Using DOLA: A Privacy-Preserving, LLM-Based Optimization AgentarXiv2025.03Paper-
GPT-PlanA Feasibility Study of Automating Radiotherapy Planning with Large Language Model AgentsPhysics in Medicine and Biology2025.03Paper-
VILA-M3VILA-M3: Enhancing Vision-Language Models with Medical Expert KnowledgeCVPR252024.11PaperProject
CT-AgentCT-Agent: A Multimodal-LLM Agent for 3D CT Radiology Question AnsweringarXiv2025.05Paper-
M^3BuilderM^3Builder: A Multi-Agent System for Automated Machine Learning in Medical ImagingAI for Clinical Applications 20252025.05Paper-
SAMIRATowards user-centered interactive medical image segmentation in VR with an assistive AI agentarXiv2025.05Paper-
MAMMAM: Modular Multi-Agent Framework for Multi-Modal Medical Diagnosis via Role-Specialized CollaborationFindings of ACL 20252025.07 (arXiv: 2025.06)PaperProject
AgentMRIAgentMRI: A Vision Language Model-Powered AI System for Self-regulating MRI Reconstruction with Multiple DegradationsJournal of Imaging Informatics in Medicine2025.07Paper-
CTPA-AgentVision-language model for report generation and outcome prediction in CT pulmonary angiogramnpj Digital Medicine2025.07PaperProject
TissueLabA co-evolving agentic AI system for medical imaging analysisarXiv2025.09PaperProject
Scan-do AttitudeScan-do Attitude: Towards Autonomous CT Protocol Management Using a Large Language Model AgentAgentic AI for Medicine / Springer2025.09Paper-
VoxelPromptVoxelPrompt: A Vision Agent for End-to-End Medical Image AnalysisarXiv2025.10Paper-
MedAgentSimMedAgentSim: Self-Evolving Multi-Agent Simulations for Realistic Clinical InteractionsMICCAI 20252025.10 (arXiv: 2025.03)PaperProject
AURAAURA: A Multi-Modal Medical Agent for Understanding, Reasoning & AnnotationMICCAI Workshop 20252025.10 (arXiv: 2025.07)PaperProject
MedEyesMedEyes: Learning Dynamic Visual Focus for Medical Progressive DiagnosisarXiv2025.11PaperProject
MedSAM3MedSAM3: Delving into Segment Anything with Medical ConceptsarXiv2025.11PaperProject
Radiologist CopilotRadiologist Copilot: An Agentic Framework Orchestrating Specialized Tools for Reliable Radiology ReportingarXiv2025.12Paper-
INFORM-CTINFORM-CT: INtegrating LLMs and VLMs FOR Incidental Findings Management in Abdominal CTMIDL 20262025.12PaperProject
IBISAgentIBISAgent: Reinforcing Pixel-Level Visual Reasoning in MLLMs for Universal Biomedical Object Referring and SegmentationarXiv2026.01Paper-
MedVistaGymMEDVISTAGYM: A Scalable Training Environment for Thinking with Medical Images via Tool-Integrated Reinforcement LearningarXiv2026.01Paper-
An Explainable Agentic AI Framework for Uncertainty-Aware and Abstention-Enabled Acute Ischemic Stroke Imaging DecisionsarXiv2026.01Paper-
LungNoduleAgentLungNoduleAgent: A Collaborative Multi-Agent System for Precision Diagnosis of Lung NodulesAAAI 20262026.02 (arXiv: 2025.11)PaperProject
3DMedAgent3DMedAgent: Unified Perception-to-Understanding for 3D Medical AnalysisarXiv2026.02PaperProject
MedSAM-AgentMedSAM-Agent: Empowering Interactive Medical Image Segmentation with Multi-turn Agentic Reinforcement LearningarXiv2026.02PaperProject
CARECARE: Towards Clinical Accountability in Multi-Modal Medical Reasoning with an Evidence-Grounded Agentic FrameworkICLR 20262026.05 (arXiv: 2026.03)PaperProject
ToolSelectPicking the Right Specialist: Attentive Neural Process-based Selection of Task-Specialized Models as Tools for Agentic Healthcare SystemsarXiv2026.02Paper-
CoMMaCoMMa: Contribution-Aware Medical Multi-Agents From A Game-Theoretic PerspectivearXiv2026.02Paper-
MedSegAgentMedSegAgent: A Universal and Scalable Multi-Agent System for Instructive Medical Image SegmentationIEEE JBHI2026.03PaperProject
CT-FlowCT-Flow: Orchestrating CT Interpretation Workflow with Model Context Protocol ServersarXiv2026.03Paper-
Agent-MIRAAgent-MIRA: AI-orchestrated Medical Imaging Agent for PET Image Retrieval and AssistanceComputerized Medical Imaging and Graphics2026.03Paper-
MeissaMeissa: Multi-modal Medical Agentic IntelligencearXiv2026.03PaperProject
BT-RADS AgentAgentic Automation of BT-RADS Scoring: End-to-End Multi-Agent System for Standardized Brain Tumor Follow-up AssessmentarXiv2026.03Paper-
TheraAgentTheraAgent: Multi-Agent Framework with Self-Evolving Memory and Evidence-Calibrated Reasoning for PET TheranosticsarXiv2026.03Paper-
MedOpenClawMEDOPENCLAW: Auditable Medical Imaging Agents Reasoning over Uncurated Full StudiesarXiv2026.03Paper-
MedMASLabMedMASLab: A Unified Orchestration Framework for Benchmarking Multimodal Medical Multi-Agent SystemsarXiv2026.03PaperProject
ClinicalAgentsClinicalAgents: Multi-Agent Orchestration for Clinical Decision Making with Dual-MemoryarXiv2026.03Paper-
Doctorina MedBenchDoctorina MedBench: End-to-End Evaluation of Agent-Based Medical AIarXiv2026.03Paper-
SEERSkill-Evolving Grounded Reasoning for Free-Text Promptable 3D Medical Image SegmentationarXiv2026.03Paper
RadAgentRadAgent: A Tool-Using AI Agent for Stepwise Interpretation of Chest Computed TomographyarXiv2026.04PaperProject
BAAI Cardiac AgentBAAI Cardiac Agent: An intelligent multimodal agent for automated reasoning and diagnosis of cardiovascular diseases from cardiac magnetic resonance imagingarXiv2026.04PaperProject
DosimeTronDosimeTron: Automating Personalized Monte Carlo Radiation Dosimetry in PET/CT with Agentic AIarXiv2026.04Paper-
MARCHMARCH: Multi-Agent Radiology Clinical Hierarchy for CT Report GenerationACL 20262026.04Paper-
Neuro-Radiological AgentAgentic Large Language Models for Training-Free Neuro-Radiological Image AnalysisarXiv2026.04Paper-
Agent4MRAgentic MR sequence development: leveraging LLMs with MR skills for automatic physics-informed sequence developmentarXiv2026.04Paper-
Artifact-based Agent FrameworkAn Artifact-based Agent Framework for Adaptive and Reproducible Medical Image ProcessingarXiv2026.04Paper-
NeuroClawNeuroClaw: Closed-Loop Agentic AI for Executable and Reproducible Neuroimaging ResearcharXiv2026.04Paper-
Neuro-OracleNeuro-Oracle: A Trajectory-Aware Agentic RAG Framework for Interpretable Epilepsy Surgical PrognosisarXiv2026.04Paper-
MedScribeMedScribe: Clinically Grounded CT Reporting through Agentic WorkflowsarXiv2026.05Paper-
GAZEGAZE: Grounded Agentic Zero-shot Evaluation with Viewer-Level Tools and Literature Retrieval on Rare Brain MRIarXiv2026.05Paper-
NeuroAgentNeuroAgent: LLM Agents for Multimodal Neuroimaging Analysis and ResearcharXiv2026.05Paper-
NEXUSTowards a Virtual Neuroscientist: Autonomous Neuroimaging Analysis via Multi-Agent CollaborationarXiv2026.05PaperProject
M2M-LLM-RTA Machine-to-Machine Knowledge-Guided LLM Agent for Generalizable Radiotherapy Treatment PlanningarXiv2026.05Paper-
SpineAgentA Multi-Agent System for Spine MRI Report Generation from Multi-Sequence ImagingarXiv2026.06Paper-
MedToolicaMedToolica: Finetuning-Free Agentic Compositional Tool Learning for 3D CT ReasoningMachine Learning and Knowledge Extraction2026.06PaperProject
MARTPMARTP: A Multi-Agent Simulation Framework for Automated Radiation Therapy Planning Based on LLMsPhysics in Medicine and Biology2026.06Paper-
SAGEAutomated Stereotactic Radiosurgery Planning Using a Human-in-the-Loop Reasoning Large Language Model AgentResearch Square2026.06Paper-
PET/CT AgentEnd-to-End PET/CT Interpretation and Quantification with an LLM-Orchestrated AI Agent: A Real-World Pilot StudyJournal of Nuclear Medicine2026.06Paper-
PD-CTAgentPolicy-Driven CT-Agent: Modeling Phase-Aware Diagnostic Control for Clinically Consistent CT ReasoningarXiv2026.07Paper-
MonteRETMonteRET: AI Agent Enhancing Multimodal LLMs with Multi-granularity Knowledge Retrieval for Chest CT Report GenerationarXiv2026.07Paper-
One-for-AllOne-for-All Adaptive Radiotherapy Planning Agent: A Foundation Framework for Daily CBCT-guided RadiotherapyarXiv2026.07Paper-

Datasets & Benchmarks

DatasetTitleDateVenuePaper LinkProject
SLIVER07Segmentation in the liver 2007 (SLIVER07) challenge2007.09MICCAI WorkshopPaperProject
SKI10Segmentation of Knee Images 2010 (SKI10)2010.09MICCAI ChallengePaperProject
LIDC-IDRIData From LIDC-IDRI: The Lung Image Database Consortium and Image Database Resource Initiative2011.06Med PhysPaperProject
LOLA11LOLA11: LObe and Lung Analysis 2011 Challenge2011.09MICCAI WorkshopProject
STACOM 2011 Motion TrackingSTACOM 2011: Cardiac Motion Tracking Challenge2011.09MICCAI WorkshopPaperProject
Mindboggle-101Mindboggle-101: Evaluating Brain Image Labeling Methods2012.09NeuroImagePaperProject
PROMISE12PROMISE12: Prostate MR Image Segmentation 2012 Challenge2012.10MICCAI ChallengePaperProject
NLSTThe National Lung Screening Trial: overview and study design2013.01RadiologyProject
Farsiu Ophthalmology 2013Quantitative Classification of Eyes with and without Intermediate Age-related Macular Degeneration Using Optical Coherence Tomography2013.03OphthalmologyPaperProject
Prostate-3TData From Prostate-3T2013.06TCIA CollectionProject
MRBrainS13MRBrainS13: Grand Challenge on MR Brain Image Segmentation2013.09MICCAI ChallengeProject
Chiu BOE 2014Kernel regression based segmentation of optical coherence tomography images with diabetic macular edema2014.01Biomed Opt ExpressPaperProject
Srinivasan BOE 2014Fully automated detection of diabetic macular edema and dry age-related macular degeneration from optical coherence tomography images2014.03Biomed Opt ExpressPaperProject
orCaScoreAn evaluation of automatic coronary artery calcium scoring methods with cardiac CT using the orCaScore framework2014.09MICCAI ChallengePaperProject
CETUS2014CETUS: Cardiac Echocardiography Tracking and Segmentation Challenge2014.09MICCAI ChallengeProject
Prostate-DiagnosisPROSTATE-DIAGNOSIS: Multiparametric MRI for Prostate Cancer2015.03TCIA CollectionProject
BTCVMICCAI multi-atlas labeling beyond the cranial vault-workshop and challenge2015.04MICCAI WorkshopProject
ISMRM2015 HARDIISMRM 2015 Tractography Challenge2015.06ISMRM ChallengeProject
NEATBrainS15NEATBrainS15: Neonatal Brain Structure Segmentation2015.09MICCAI ChallengeProject
PDDCAPublic Domain Database for Computational Anatomy: Head and Neck2015.09MICCAI ChallengeProject
HVSMR 2016HVSMR 2016: Whole Heart and Great Vessel Segmentation Challenge2016.07MICCAI ChallengePaperProject
MSSEG 2016Objective Evaluation of Multiple Sclerosis Lesion Segmentation using a Data Management and Processing Infrastructure2016.10MICCAI Challenge/NaturePaperProject
PROSTATExPROSTATEx: PROSTATE MR Image Dataset With Prostate Cancer Annotations2016.10SPIE-AAPM-NCI PROSTATEx ChallengeProject
WMHWMH Segmentation Challenge: White Matter Hyperintensity Segmentation in Brain MR2017.03MICCAI ChallengePaperProject
LGG-1p19qDeletionLGG-1p19qDeletion: Low-Grade Glioma MRI with Genomic Annotations2017.03TCIA CollectionProject
PROSTATEx-2PROSTATEx-2: Lesion Classification Challenge2017.06AAPM Grand ChallengeProject
ACDCAutomatic Cardiac Diagnosis Challenge2017.09STACOM / MICCAIPaperProject
RETOUCHRETOUCH: Retinal OCT Fluid Segmentation Challenge2017.09MICCAI ChallengePaperProject
ROCCROCC: Retinal OCT Classification Challenge2017.09MICCAI ChallengeProject
iSeg2017iSeg-2017: Infant Brain MRI Segmentation Challenge2017.09MICCAI ChallengePaperProject
DeepLesionDeepLesion: automated mining of large-scale lesion annotations and universal lesion detection in CT2017.10JMI / arXivPaperProject
LUNA 16Validation, comparison, and combination of algorithms for automatic detection of pulmonary nodules in computed tomography images: the LUNA 16 challenge2017.12Medical Image AnalysisPaperProject
Mandibular-CT-DatasetMandibular CT Dataset Collection for 3D Reconstruction and Segmentation2018.03figsharePaperProject
FUMPEComputer-aided detection of pulmonary embolism in CT2018.03arXiv / KaggleProject
ISLES 2018ISLES 2018 – Ischemic Stroke Lesion Segmentation2018.09MICCAI ChallengePaperProject
MRBrainS18MRBrainS18: MR Brain Segmentation Challenge 20182018.09MICCAI ChallengeProject
Atrial Segmentation Challenge2018 Atrial Segmentation Challenge2018.09MICCAI ChallengeProject
IVDM3SegIVDM3Seg: Intervertebral Disc and Vertebrae Segmentation Challenge2018.09MICCAI ChallengeProject
MRNetMRNet: Knee MRI Dataset for Abnormality Detection2018.09NIPS WorkshopProject
OCT Glaucoma DetectionGlaucoma Detection in 3D Spectral-Domain OCT2018.10Sci RepProject
BraTSBrain Tumor Segmentation (BraTS) Challenge2018.11MICCAI Challenge (series)PaperProject
fastMRIfastMRI: A Publicly Available Raw k-Space and DICOM Dataset of Knee and Brain MR Images2018.11MRMPaperProject
LiTSLiver Tumor Segmentation (LiTS) Challenge2019.01MICCAI ChallengePaperProject
OASIS-3OASIS-3: Longitudinal Neuroimaging, Clinical, and Cognitive Dataset for Normal Aging and Alzheimer Disease2019.01Sci DataPaperProject
MM-WHSMM-WHS: Multi-Modality Whole Heart Segmentation2019.02MICCAI ChallengePaperProject
CHAOS CT-MRICHAOS - Combined (CT-MR) Healthy Abdominal Organ Segmentation2019.02ISBI ChallengePaperProject
KiTS19KiTS19: Kidney Tumor Segmentation Challenge2019.04MICCAI ChallengePaperProject
AAPM-RT-MACAAPM RT-MAC: MR-only based Radiotherapy in Head and Neck2019.07AAPM ChallengePaperProject
iSeg-2019iSeg-2019: Infant Brain MRI Segmentation Challenge2019.09MICCAI ChallengePaperProject
SegTHORSegTHOR: Segmentation of thoracic organs at risk in CT images2019.09Physica MedicaPaperProject
VerSe20VerSe 2020: Vertebral Segmentation Challenge at MICCAI2020.01MICCAI ChallengeProject
VerSe19VerSe 2019: Vertebral Segmentation Challenge at MICCAI2020.01MICCAI ChallengePaperProject
COVID-19-CT-SegCOVID-19 CT lung and infection segmentation dataset2020.04zenodoProject
M&MsM&Ms: Multi-Centre, Multi-Vendor & Multi-Disease Cardiac MR Segmentation Challenge2020.05MICCAI ChallengeProject
CTPelvic1KCTPelvic1K: A Large-Scale Pelvic CT Dataset for Multi-Task Parsing2020.06arXivPaperProject
Prostate MR Segmentation Dataset (SAML)Federated Domain Generalization on Medical Image Segmentation via Episodic Learning in Continuous Frequency Space2020.09MICCAIProject
EMIDECEMIDEC 2020: Myocardial Infarction Detection, Segmentation and Classification2020.09MICCAI ChallengePaperProject
KNOAP2020KNOAP2020: Knee Osteoarthritis Progression Prediction Challenge2020.09MICCAI ChallengeProject
Learn2Reg Lung CTLearn2Reg 2020: Lung CT Registration2020.09MICCAI ChallengeProject
Learn2Reg Abdomen CT-CTLearn2Reg 2020: Abdominal CT-CT Registration2020.09MICCAI ChallengeProject
RibFrac2020RibFrac: Rib Fracture Detection and Classification Challenge2020.10MICCAI ChallengeProject
HECKTOR 2020HECKTOR 2020: Segmentation of Head and Neck Tumor in PET/CT2020.11MICCAI ChallengeProject
CT-ORGCT-ORG, a new dataset for multiple organ segmentation in computed tomography2020.11NaturePaperProject
RAD-ChestCTMachine-Learning-Based Multiple Abnormality Prediction with Large-Scale Chest Computed Tomography Volumes2021.01Medical Image AnalysisPaperProject
Eye OCT Datasets (3D)3D Retinal OCT Classification and Segmentation Dataset2021.01TianchiProject
HECKTOR 2021HECKTOR 2021: Head and Neck Tumor Segmentation and Outcome Prediction2021.05MICCAI ChallengeProject
CTSpine1KCTSpine1K: A Large-Scale Dataset for Spine Parsing in CT2021.07arXivProject
MSSEG-2MSSEG-2 challenge: Multiple Sclerosis Lesion Segmentation at 7T and 3T MRI2021.07NeuroImage ClinProject
FLARE21FLARE 2021: A Challenge on Abdominal Multi-organ Segmentation2021.09MICCAI ChallengeProject
QUBIQ2021 3D CTQUBIQ 2021: Quantification of Uncertainty in Biomedical Image Quantification2021.09MICCAI ChallengeProject
M&Ms-2M&Ms-2: Multi-Domain Cardiac MR Segmentation2021.09MICCAI ChallengeProject
Learn2Reg Abdomen MR-CTLearn2Reg 2021: Abdominal MR-CT Multi-Modal Registration2021.09MICCAI ChallengeProject
CrossMoDA2021CrossMoDA 2021: Unsupervised Domain Adaptation for Cross-Modality Vestibular Schwannoma Segmentation2021.09MICCAI ChallengeProject
WORDWORD: A Whole-Organ CT Dataset for Robust Multi-Organ Segmentation2021.10arXivPaperProject
MedMNIST v2MedMNIST v2 -- A large-scale lightweight benchmark for 2D and 3D biomedical image classification2021.10NaturePaperProject
CADACADA: Cerebral Aneurysm Detection and Analysis Challenge2022.04MICCAI ChallengeProject
CADA-ASCADA-AS: Aneurysm Segmentation Challenge2022.04MICCAI ChallengeProject
CADA-RRECADA-RRE: Rupture Risk Estimation for Cerebral Aneurysms2022.04MICCAI ChallengeProject
TotalSegmentatorTotalSegmentator: robust segmentation of 104 anatomic structures in CT images2022.06arXivPaperProject
AMOSAMOS: A Large-Scale Abdominal Multi-Organ Benchmark for Versatile Medical Image Segmentation2022.06NeurIPS 2022PaperProject
MSDThe medical segmentation decathlon2022.07Nature CommunicationsPaperProject
AutoPETThe AutoPET Challenge: Automated Lesion Segmentation in Whole-Body FDG-PET/CT2022.07MICCAI Challenge (autoPET2022)Project
UPENN-GBMUPENN-GBM: Multi-modal MRI Dataset for Glioblastoma Segmentation2022.07TCIA CollectionProject
KiPA22KiPA22: Kidney PArametric segmentation in contrast-enhanced CT2022.08MICCAI ChallengeProject
OLIVESOLIVES: A 3D OCT Dataset for Longitudinal Retinal Imaging2022.09arXivProject
PI-CAIPI-CAI: Prostate Imaging–Cancer AI Challenge2022.09MICCAI ChallengeProject
LAScarQS 2022LAScarQS 2022: Left Atrial Scar Quantification and Segmentation Challenge2022.09MICCAI ChallengeProject
CrossMoDA2022CrossMoDA 2022: Domain Adaptation for Vestibular Schwannoma Segmentation and Koos Grading2022.09MICCAI ChallengeProject
FeTA 2022FeTA 2022: Fetal Brain Tissue Segmentation at MICCAI2022.09MICCAI ChallengeProject
COSMOS 2022COSMOS 2022: Carotid Artery Vessel Wall Segmentation2022.09MICCAI ChallengeProject
cSeg-2022cSeg 2022: Cerebellum Segmentation Challenge2022.09MICCAI ChallengeProject
ISLES 2022ISLES 2022 – Acute and Subacute Ischemic Stroke Lesion Segmentation2022.09MICCAI ChallengeProject
InSTANCE2022InSTANCE 2022: Intracranial Hemorrhage Segmentation Challenge2022.09MICCAI ChallengePaperProject
Learn2Reg NLSTLearn2Reg 2022: Thoracic CT Registration with NLST2022.09MICCAI ChallengeProject
Shifts Challenge 2022Shifts 2022: Distribution Shifts in Multiple Sclerosis Lesion Segmentation2022.09MICCAI ChallengePaperProject
HECKTOR 22Overview of the HECKTOR challenge at MICCAI 2022: automatic head and neck tumor segmentation and outcome prediction in PET/CT2023MICCAI 2022PaperProject
LNDbLNDb challenge on automatic lung cancer patient management2023.03Medical Image AnalysisPaperProject
Semi-TeethSegSemi-TeethSeg: Semi-Supervised 3D Tooth Segmentation in CBCT/CT2023.04arXivProject
PARSE22Efficient automatic segmentation for multi-level pulmonary arteries: The parse challenge2023.04arXivPaperProject
STAGESTAGE: Longitudinal OCT Dataset for Glaucoma Progression2023.04DatasetProject
KiTS21The kits21 challenge: Automatic segmentation of kidneys, renal tumors, and renal cysts in corticomedullary-phase ct2023.07arXivPaperProject
AutoPET IIAutoPET-II: Ensemble-based Uncertainty-Aware Lesion Segmentation in Multi-Center FDG-PET/CT2023.07MICCAI Challenge (AutoPET-II)Project
MedMDTowards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical Data2023.08arXivPaperProject
ULS23ULS23 Challenge: Universal Lesion Segmentation in CT for Oncological Imaging2023.08MICCAI ChallengeProject
SegRap2023SegRap 2023: Nasopharyngeal Carcinoma Radiotherapy Segmentation Challenge2023.08MICCAI ChallengeProject
LNQ2023LNQ2023: Lymph Node Quantification in Chest CT2023.08MICCAI ChallengeProject
FLARE23FLARE 2023: A Federated Learning Challenge for Abdominal Multi-Organ Segmentation2023.09MICCAI ChallengeProject
CrossMoDA2023CrossMoDA 2023: Multi-Center Domain Adaptation for VS Segmentation2023.09MICCAI ChallengeProject
ATLAS2023ATLAS 2023: Liver Tumor Segmentation Challenge2023.09MICCAI ChallengePaperProject
SMILE-UHURA2023SMILE-UHURA 2023: Small Vessel Disease Lesion Segmentation2023.09MICCAI ChallengePaperProject
CAS2023CAS 2023: Brain Structure Segmentation Benchmark2023.09MICCAI ChallengeProject
CROWN2023CROWN 2023: White Matter Hyperintensity and Other Pathology Classification2023.09MICCAI ChallengeProject
SLCNSLCN: Structural Lesion and Connectivity in Neurodevelopmental Disorders2023.09MICCAI ChallengeProject
ToothFairy2023ToothFairy: 3D CBCT Dataset for Inferior Alveolar Nerve Segmentation2023.09MICCAI ChallengeProject
XPRESS2023XPRESS 2023: X-ray Phase-Contrast CT Neuroanatomy Segmentation2023.09MICCAI ChallengePaperProject
Learn2Reg ThoraxCBCTLearn2Reg 2023: Thorax CBCT/FBCT Deformable Registration2023.09MICCAI ChallengeProject
TDSC-ABUS2023TDSC-ABUS2023: Automated Breast Ultrasound Segmentation Challenge2023.09MICCAI ChallengePaperProject
MVSeg-3DTEE2023MVSeg-3DTEE2023: Mitral Valve Segmentation from 3D TEE2023.09MICCAI ChallengeProject
RegPro2023RegPro 2023: Prostate MR-US Registration Challenge2023.09MICCAI ChallengeProject
KiTS23KiTS23: Kidney and Kidney Tumor Segmentation with Comprehensive Clinical Annotations2023.10arXivProject
WBMR-NFWBMR-NF: Whole-Body MRI for Neurofibromatosis2023.11DatasetProject
GAMMAGAMMA Challenge: Glaucoma Assessment with Multi-Modality Data2023.12MICCAI ChallengePaperProject
ATM'22ATM'22: Airway Tree Modeling in Thoracic CT2023.12MICCAI ChallengePaperProject
INSPECTINSPECT: A Multimodal Dataset for Pulmonary Embolism Diagnosis and Prognosis2023.12NeurIPS 2023PaperProject
HaN-SegHaN-Seg 2023: Head and Neck Organ at Risk Segmentation Challenge2024MICCAI ChallengePaperProject
VALDOWhere is VALDO? Vascular Lesions Detection and Segmentation Challenge2024.01MICCAI ChallengePaperProject
BIMCV-RBIMCV-R: Large-Scale Thoracic CT Reconstruction Benchmark2024.01MICCAI24PaperProject
IXIInformation eXtraction from Images (IXI) Dataset2024.01DatasetProject
RAOSRAOS: A Large-Scale Radiotherapy Abdominal Organ Segmentation Dataset2024.01MICCAI24PaperProject
ISLES 2024Ischemic Stroke Lesion Segmentation Challenge 2024 (ISLES 2024)2024.02MICCAI ChallengeProject
AbdomenAtlasAbdomenAtlas-20K: A Large-Scale Benchmark for Abdominal Multi-Organ Segmentation in CT2024.02arXivPaperProject
OpenMindOpenMind: Large-Scale Head-and-Neck MR Dataset for Foundation Models2024.02arXivProject
TriALS2024TriALS 2024: Liver Tumor Segmentation and Outcome Prediction – Task 12024.03MICCAI24Project
LAScarQS++ 2024LAScarQS++ 2024: Multi-Center Atrial Scar Segmentation2024.03CARE Workshop (MICCAI)Project
MyoPSMyoPS: A Benchmark of Myocardial Pathology Segmentation Combining Three-Sequence Cardiac Magnetic Resonance Images# MyoPS: A Benchmark of Myocardial Pathology Segmentation Combining Three-Sequence Cardiac Magnetic Resonance Images2024.03CARE WorkshopPaperProject
WHS++ 2024WHS++ 2024: Multi-Center Whole Heart Segmentation2024.03CARE WorkshopProject
AMOS-MMAMOS-MM: Multi-Phase Abdominal CT Benchmark for Translation and Synthesis2024.03arXivPaperProject
CT2RepCT2Rep: Automated Radiology Report Generation for 3D Medical Imaging2024.03MICCAI 2024PaperProject
TotalSegmentator MRITotalSegmentator MRI: Whole-body MRI Segmentation of 150 Structures2024.04arXivPaperProject
RadGenome-ChestCTRadGenome-Chest CT: a grounded vision-language dataset for chest CT analysis2024.04arXivPaperProject
M3DM3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models2024.04ICLR 2025PaperProject
AIIB23AIIB23: Airway Inflammation Imaging Biomarkers Challenge2024.06MICCAI ChallengePaperProject
CT-3DRRGArgus: Benchmarking and Enhancing Vision-Language Models for 3D Radiology Report Generation2024.06arXivPaper—
RadGenome-Brain MRIAutoRG-Brain: Grounded Report Generation for Brain MRI2024.10MICCAI 2024PaperProject
MedShapeNetMedShapeNet -- A Large-Scale Dataset of 3D Medical Shapes for Computer Vision2024.12Biomedizinische TechnikPaperProject
RadA-BenchPlatHow well can modern LLMs act as agent cores in radiology environments?2024.12arXivPaper
MedVL-CT69KLarge-scale and Fine-grained Vision-language Pre-training for Enhanced CT Image Understanding2025.01ICLR 2025PaperProject
TriadTriad: Vision Foundation Model for 3D Magnetic Resonance Imaging2025.02arXivPaper
3D-BrainCTTowards a holistic framework for multimodal LLM in 3D brain CT radiology report generation2025.03Nat. Commun.Paper—
PENGWIN2024-Task1PENGWIN 2024: Pelvic Fracture Segmentation in Trauma CT2025.04MICCAI ChallengePaperProject
RibFracDeep rib fracture instance segmentation and classification from ct on the ribfrac challenge2025.04IEEEPaperProject
DeepTumorVQAAre Vision Language Models Ready for Clinical Diagnosis? A 3D Medical Benchmark for Tumor-centric Visual Question Answering2025.05NeurIPS 2025PaperProject
NOVANOVA: A Benchmark for Anomaly Localization and Clinical Reasoning in Brain MRI2025.05NeurIPS 2025Paper
LingshuLingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning2025.06arXivPaperProject
ReXGroundingCTReXGroundingCT: A 3D Chest CT Dataset for Segmentation of Findings from Free-Text Reports2025.07arXivPaperProject
ViPET-ReportGenToward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation2025.12NeurIPS 2025PaperProject
3D-RAD3D-RAD: A Comprehensive 3D Radiology Med-VQA Dataset with Multi-Temporal Analysis and Diverse Diagnostic Tasks2025.12NeurIPS 2025PaperProject
MR-RATEMR-RATE: A Vision-Language Foundation Model and Dataset for Magnetic Resonance Imaging2026——Project
CT-RATEGeneralist Foundation Models from a Multimodal Dataset for 3D Computed Tomography2026.02NaturePaperProject
CT-FlowBenchCT-FlowBench: Benchmark for CT interpretation workflow and tool-use2026.03arXivPaper
MerlinMerlin: A Vision Language Foundation Model for 3D Computed Tomography2026.03NaturePaperProject
Gastric-XGastric-X: A Multimodal Multi-Phase Benchmark Dataset for Advancing Vision-Language Models in Gastric Cancer Analysis2026.03CVIPPR 2026Paper—
SpatialMedBeyond Medical Diagnostics: How Medical Multimodal Large Language Models Think in Space2026.03arXivPaper
BONBID-HIE2023BONBID-HIE 2023: Neonatal Hypoxic-Ischemic Encephalopathy Lesion Segmentation2026.04MICCAI ChallengePaperProject
SGMRI-VQABeyond a Single Frame: Multi-Frame Spatially Grounded Reasoning Across Volumetric MRI2026.04arXivPaper
Curia-2Curia-2: Scaling Self-Supervised Learning for Radiology Foundation Models2026.04arXivPaper
CT-SpatialVQALost in Volume: The CT-SpatialVQA Benchmark for Evaluating Semantic-Spatial Understanding of 3D Medical Vision-Language Models2026.05arXivPaper
Med-StepBenchMed-StepBench: A Hierarchical Reasoning Framework for Evaluating Hallucinations in Medical Vision-Language Models2026.05arXivPaper
DeepTumorVQA-HDeepTumorVQA: A Hierarchical 3D CT Benchmark for Stage-Wise Evaluation of Medical VLMs and Tool-Augmented Agents2026.05arXivPaper
ABRAABRA: Agent Benchmark for Radiology Applications2026.05arXivPaper
RadSaFE-200Safety and Accuracy Follow Different Scaling Laws in Clinical Large Language Models2026.05arXivPaper
Oncology VQA BenchmarkAutomated Report-Derived Oncology VQA Benchmark for Evaluating Vision-Language Models on 3D Medical Imaging2026.06arXivPaper
Abdomen-NCCT BenchmarkA Multi-Center Benchmark for Abdominal Disease Diagnosis and Report Generation from Non-Contrast CT2026.06arXivPaper
RadOT-EvalRadOT-Eval: Auditable Structured-Evidence Transport for Radiology Report Evaluation2026.06arXivPaper
ReportQAReportQA: QA-Based Radiology Report Evaluation2026.06arXivPaper
CORTEXCORTEX: A Structured Reasoning Benchmark for Trustworthy 3D Chest CT MLLMs2026.06arXivPaper
MedCTAMedCTA: A Benchmark for Clinical Tool Agents2026.06arXivPaper
Lung CT FM BenchmarkFoundation Models vs. Radiomics for Lung Computed Tomography: A Benchmark of Feature Extractors, Classification Heads, and Segmentation Choices2026.07arXivPaperProject
Brain Oncology 3D MRI-TextMulti-LLM Collaborative MRI Report Generation for Visual Instruction Tuning in Brain Oncology2026.07arXivPaper-
COBRA2026COBRA2026: a large-scale multicenter pelvic cone-beam computed tomography projection dataset2026.07arXivPaperProject
GLI-ALGLI-AL: A Multi-Modal Glioma MRI Label Resource with Unified Anatomy-Lesion Labels2026.07arXivPaperProject

Contributors

Saint-lsy

16 commits

yezanting

13 commits

Leo-lxin

10 commits