Nanboy-Ronan/awesome-medical-imaging-agents

Awesome list for medical imaging agents: radiology agents, pathology agents, segmentation agents, medical VLM agents, self-evolving agents, benchmarks, datasets, tools, and papers.

Python

54

101 commits

updated Sep 23, 2026

See the code

README

Awesome Medical Imaging Agents Awesome

Agentic AI systems for medical image analysis, including radiology agents, pathology agents, ultrasound agents, surgical imaging agents, segmentation agents, and medical vision-language model agents.

This repository curates research papers, benchmarks, datasets, and open-source systems for medical imaging agents. It focuses on tool use, retrieval-augmented generation, multi-agent collaboration, planning, segmentation, report generation, clinical reasoning, safety, and evaluation across CT, MRI, chest X-ray, ultrasound, pathology, endoscopy, ophthalmology, PET, and other imaging modalities.

Taxonomy

Visual taxonomy of medical imaging agents

Start Here

Twelve landmark systems, one per major domain, for readers who want the fastest path into the field.

PaperDomainYearWhy it mattersCode
MedRAX: Medical Reasoning Agent for Chest X-rayRadiology2025Director-worker architecture where composable tool-using agents outperform single-model baselines on chest X-ray reasoning.Code
RadAgent: A Tool-Using AI Agent for Stepwise Interpretation of Chest CTRadiology2026Generates chest CT reports through an explicit tool-calling workflow with inspectable intermediate reasoning traces.—
Agentic Systems in Radiology: Design, Applications, Evaluation, and ChallengesSurvey · Radiology2025Best entry-point survey mapping agent design patterns, evaluation protocols, and open challenges across the full radiology pipeline.—
CPathAgent: An Agent-based Foundation Model for Interpretable High-Resolution Pathology Image AnalysisPathology2025Agentic pathology foundation model mimics pathologist diagnostic logic to deliver interpretable analysis of whole-slide images.—
Echo-alpha: Large Agentic Multimodal Reasoning Model for Ultrasound InterpretationUltrasound2026Demonstrates that large agentic reasoning models can jointly localize lesions and perform grounded multimodal clinical interpretation of ultrasound studies.—
EndoAgent: A Memory-Guided Reflective Agent for Intelligent Endoscopic Vision-to-Decision ReasoningEndoscopy2025Memory-guided reflective agent for endoscopic vision-to-decision reasoning; representative model for specialty-specific imaging agent design.Code
VoxelPrompt: A Vision Agent for End-to-End Medical Image Analysis3D Imaging2024Multi-stage vision agent for end-to-end volumetric medical image analysis covering segmentation, detection, and QA across CT, MRI, and PET.—
MedSAM-Agent: Empowering Interactive Medical Image Segmentation with Multi-turn Agentic Reinforcement LearningSegmentation2026Introduces multi-turn RL to train an interactive segmentation agent that decides which prompts to issue to SAM — the first RL-trained imaging segmentation agent.Code
MMedAgent: Learning to Use Medical Tools with Multi-modal AgentMultimodal2024First paper to train a medical agent that selects and calls specialist tools (segmentation, retrieval, calculators) on demand across seven imaging modalities.Code
MedAgent-Pro: Towards Evidence-based Multi-modal Medical Diagnosis via Reasoning Agentic WorkflowMedical VLM2026Integrates imaging, labs, and clinical guidelines via explicit tool calling; demonstrates evidence-based multimodal clinical reasoning.Code
AgentClinic: A Multimodal Benchmark for Tool-Using Clinical AI AgentsBenchmark2026The canonical multimodal benchmark for tool-using clinical AI agents with an open simulator across imaging, EHR, and lab modalities.Code
ABRA: Agent Benchmark for Radiology ApplicationsBenchmark · Radiology2026First benchmark where agents operate a real DICOM viewer (OHIF + Orthanc) via tool calls, testing end-to-end radiology agent workflows on live imaging software.—

Scope

  • Medical imaging agents for radiology, pathology, ultrasound, CT, MRI, chest X-ray, dermatology, endoscopy, PET, ophthalmology, and cardiac imaging.
  • Agent workflows for tool use, retrieval-augmented generation, multi-agent collaboration, self-reflection, planning, report generation, quality control, and human-in-the-loop review.
  • Research resources including peer-reviewed papers, arXiv preprints, benchmarks, datasets, reproducible code, and related open-source systems.
  • Safety and evaluation work on hallucination detection, fairness, robustness, uncertainty, abstention, privacy, and clinically grounded evaluation.

Contents

Radiology Agents (44)

Agents for chest X-ray, CT, MRI, DICOM workflows, radiotherapy planning, and radiology decision support.

Pathology Agents (Whole-Slide Imaging · Digital Pathology) (31)

Agents for whole-slide image analysis, digital pathology, pathology reports, and slide navigation.

Ultrasound Agents (Echocardiography · Robotic Ultrasound) (20)

Agents for echocardiography interpretation, fetal ultrasound, robotic scanning, and ultrasound-guided workflows.

Endoscopy and Surgical Imaging Agents (18)

Agents for gastrointestinal endoscopy, surgical scene understanding, and autonomous endoscopic navigation.

Ophthalmology Agents (13)

Agents for fundus, OCT, glaucoma, diabetic retinopathy, myopia, and neuro-ophthalmic decision support.

3D CT / MRI / Volumetric Imaging Agents (21)

Agents for volumetric CT, MRI, PET, dosimetry, neuroimaging, and multi-organ image analysis.

Segmentation and Annotation Agents (14)

Agents that plan, prompt, refine, or evaluate segmentation and annotation workflows.

Report Generation Agents (17)

Agents focused on automated imaging report drafting, refinement, evaluation, and quality control.

Medical Vision-Language Model (VLM) Agents (30)

Vision-language agents that combine imaging encoders, language models, tools, and clinical reasoning. Includes broad multimodal agents that reason jointly over multiple imaging modalities and clinical text.

Backbone Foundation Models (not agents) (32)

Pretrained medical LLMs, multimodal LLMs, and image encoders frequently wrapped by the agent systems above; included for reference, not as agents themselves.

Tool-Using and Multi-Agent Frameworks

Agents and frameworks for general clinical reasoning, workflow automation, simulation, and tool/skill learning that span beyond a single imaging modality.

Clinical Reasoning Agents (76)

Agents for diagnosis, differential reasoning, treatment planning, retrieval, and clinical decision support.

Workflow and Simulation Agents (49)

Agents and environments for clinical workflow automation, simulation, and operational task execution.

Agent Skills and Tool Learning (10)

Agents and audits focused specifically on how medical agents acquire, retrieve, and govern reusable tools and skills.

Benchmarks and Evaluation

Benchmarks, datasets, simulators, and evaluation frameworks for medical and imaging agents.

Benchmark Table

Benchmarks with explicit imaging modality and task metadata.

BenchmarkModalityTaskLink
AutoMedBenchCT, MRI, CXRbenchmark, segmentation, report generation, VQAPaper
AgentClinicCXR, EHR, lab tablestool use, diagnosis, multimodal reasoningPaper · Site · Code
ABRACT, MRI, DICOMDICOM viewer navigation, tool use, report generationPaper
MedAgentBenchEHRlongitudinal task completion, clinical decision makingPaper · Code
MedAgentBoardCXR, EHRmulti-agent collaboration, diagnosis, question answeringPaper
DeepTumorVQACTvisual question answering, tool use, localizationPaper
MedRCubeCXR, CT, MRI, histopathologyevaluation, benchmarkPaper
Trustworthy Medical Imaging with LLMsCXR, CT, MRIsafety, hallucination detectionPaper
MedCTACXR, WSItool use, benchmarkPaper
Colon-Benchcolonoscopylesion detection, annotation, benchmarkPaper
MEDVISTAGYMmulti-modalitytraining environment, tool use, visual reasoningPaper
DALPHINWSIVQA, evaluation, benchmarkPaper · Site
SpatialMedCTVQA, 3D spatial reasoning, benchmarkPaper

Benchmark Papers (44)

Safety, Robustness, and Fairness (34)

Themes Index

Cross-cutting topics that span multiple domain sections above. Each paper is listed once in its primary section; this index lets you find it by theme.

Fairness and Bias (7)

Hallucination and Reliability (39)

Safety and Robustness (31)

Privacy and Federated Learning (6)

RAG and Retrieval (52)

Multi-Agent Collaboration (145)

Truncated — view the full README on GitHub.

agents
ai-agents
awesome
awesome-list
computer-vision
deep-learning
healthcare
medical-image-analysis
medical-imaging
radiology
self-evolving
self-evolving-agents
trustworth-ai

Contributors

Nanboy-Ronan

100 commits

gexinh

1 commits

Nanboy-Ronan/awesome-medical-imaging-agents

Awesome list for medical imaging agents: radiology agents, pathology agents, segmentation agents, medical VLM agents, self-evolving agents, benchmarks, datasets, tools, and papers.

Python

54

101 commits

updated Sep 23, 2026

See the code

README

Awesome Medical Imaging Agents Awesome

Agentic AI systems for medical image analysis, including radiology agents, pathology agents, ultrasound agents, surgical imaging agents, segmentation agents, and medical vision-language model agents.

This repository curates research papers, benchmarks, datasets, and open-source systems for medical imaging agents. It focuses on tool use, retrieval-augmented generation, multi-agent collaboration, planning, segmentation, report generation, clinical reasoning, safety, and evaluation across CT, MRI, chest X-ray, ultrasound, pathology, endoscopy, ophthalmology, PET, and other imaging modalities.

Taxonomy

Visual taxonomy of medical imaging agents

Start Here

Twelve landmark systems, one per major domain, for readers who want the fastest path into the field.

PaperDomainYearWhy it mattersCode
MedRAX: Medical Reasoning Agent for Chest X-rayRadiology2025Director-worker architecture where composable tool-using agents outperform single-model baselines on chest X-ray reasoning.Code
RadAgent: A Tool-Using AI Agent for Stepwise Interpretation of Chest CTRadiology2026Generates chest CT reports through an explicit tool-calling workflow with inspectable intermediate reasoning traces.—
Agentic Systems in Radiology: Design, Applications, Evaluation, and ChallengesSurvey · Radiology2025Best entry-point survey mapping agent design patterns, evaluation protocols, and open challenges across the full radiology pipeline.—
CPathAgent: An Agent-based Foundation Model for Interpretable High-Resolution Pathology Image AnalysisPathology2025Agentic pathology foundation model mimics pathologist diagnostic logic to deliver interpretable analysis of whole-slide images.—
Echo-alpha: Large Agentic Multimodal Reasoning Model for Ultrasound InterpretationUltrasound2026Demonstrates that large agentic reasoning models can jointly localize lesions and perform grounded multimodal clinical interpretation of ultrasound studies.—
EndoAgent: A Memory-Guided Reflective Agent for Intelligent Endoscopic Vision-to-Decision ReasoningEndoscopy2025Memory-guided reflective agent for endoscopic vision-to-decision reasoning; representative model for specialty-specific imaging agent design.Code
VoxelPrompt: A Vision Agent for End-to-End Medical Image Analysis3D Imaging2024Multi-stage vision agent for end-to-end volumetric medical image analysis covering segmentation, detection, and QA across CT, MRI, and PET.—
MedSAM-Agent: Empowering Interactive Medical Image Segmentation with Multi-turn Agentic Reinforcement LearningSegmentation2026Introduces multi-turn RL to train an interactive segmentation agent that decides which prompts to issue to SAM — the first RL-trained imaging segmentation agent.Code
MMedAgent: Learning to Use Medical Tools with Multi-modal AgentMultimodal2024First paper to train a medical agent that selects and calls specialist tools (segmentation, retrieval, calculators) on demand across seven imaging modalities.Code
MedAgent-Pro: Towards Evidence-based Multi-modal Medical Diagnosis via Reasoning Agentic WorkflowMedical VLM2026Integrates imaging, labs, and clinical guidelines via explicit tool calling; demonstrates evidence-based multimodal clinical reasoning.Code
AgentClinic: A Multimodal Benchmark for Tool-Using Clinical AI AgentsBenchmark2026The canonical multimodal benchmark for tool-using clinical AI agents with an open simulator across imaging, EHR, and lab modalities.Code
ABRA: Agent Benchmark for Radiology ApplicationsBenchmark · Radiology2026First benchmark where agents operate a real DICOM viewer (OHIF + Orthanc) via tool calls, testing end-to-end radiology agent workflows on live imaging software.—

Scope

  • Medical imaging agents for radiology, pathology, ultrasound, CT, MRI, chest X-ray, dermatology, endoscopy, PET, ophthalmology, and cardiac imaging.
  • Agent workflows for tool use, retrieval-augmented generation, multi-agent collaboration, self-reflection, planning, report generation, quality control, and human-in-the-loop review.
  • Research resources including peer-reviewed papers, arXiv preprints, benchmarks, datasets, reproducible code, and related open-source systems.
  • Safety and evaluation work on hallucination detection, fairness, robustness, uncertainty, abstention, privacy, and clinically grounded evaluation.

Contents

Radiology Agents (44)

Agents for chest X-ray, CT, MRI, DICOM workflows, radiotherapy planning, and radiology decision support.

Pathology Agents (Whole-Slide Imaging · Digital Pathology) (31)

Agents for whole-slide image analysis, digital pathology, pathology reports, and slide navigation.

Ultrasound Agents (Echocardiography · Robotic Ultrasound) (20)

Agents for echocardiography interpretation, fetal ultrasound, robotic scanning, and ultrasound-guided workflows.

Endoscopy and Surgical Imaging Agents (18)

Agents for gastrointestinal endoscopy, surgical scene understanding, and autonomous endoscopic navigation.

Ophthalmology Agents (13)

Agents for fundus, OCT, glaucoma, diabetic retinopathy, myopia, and neuro-ophthalmic decision support.

3D CT / MRI / Volumetric Imaging Agents (21)

Agents for volumetric CT, MRI, PET, dosimetry, neuroimaging, and multi-organ image analysis.

Segmentation and Annotation Agents (14)

Agents that plan, prompt, refine, or evaluate segmentation and annotation workflows.

Report Generation Agents (17)

Agents focused on automated imaging report drafting, refinement, evaluation, and quality control.

Medical Vision-Language Model (VLM) Agents (30)

Vision-language agents that combine imaging encoders, language models, tools, and clinical reasoning. Includes broad multimodal agents that reason jointly over multiple imaging modalities and clinical text.

Backbone Foundation Models (not agents) (32)

Pretrained medical LLMs, multimodal LLMs, and image encoders frequently wrapped by the agent systems above; included for reference, not as agents themselves.

Tool-Using and Multi-Agent Frameworks

Agents and frameworks for general clinical reasoning, workflow automation, simulation, and tool/skill learning that span beyond a single imaging modality.

Clinical Reasoning Agents (76)

Agents for diagnosis, differential reasoning, treatment planning, retrieval, and clinical decision support.

Workflow and Simulation Agents (49)

Agents and environments for clinical workflow automation, simulation, and operational task execution.

Agent Skills and Tool Learning (10)

Agents and audits focused specifically on how medical agents acquire, retrieve, and govern reusable tools and skills.

Benchmarks and Evaluation

Benchmarks, datasets, simulators, and evaluation frameworks for medical and imaging agents.

Benchmark Table

Benchmarks with explicit imaging modality and task metadata.

BenchmarkModalityTaskLink
AutoMedBenchCT, MRI, CXRbenchmark, segmentation, report generation, VQAPaper
AgentClinicCXR, EHR, lab tablestool use, diagnosis, multimodal reasoningPaper · Site · Code
ABRACT, MRI, DICOMDICOM viewer navigation, tool use, report generationPaper
MedAgentBenchEHRlongitudinal task completion, clinical decision makingPaper · Code
MedAgentBoardCXR, EHRmulti-agent collaboration, diagnosis, question answeringPaper
DeepTumorVQACTvisual question answering, tool use, localizationPaper
MedRCubeCXR, CT, MRI, histopathologyevaluation, benchmarkPaper
Trustworthy Medical Imaging with LLMsCXR, CT, MRIsafety, hallucination detectionPaper
MedCTACXR, WSItool use, benchmarkPaper
Colon-Benchcolonoscopylesion detection, annotation, benchmarkPaper
MEDVISTAGYMmulti-modalitytraining environment, tool use, visual reasoningPaper
DALPHINWSIVQA, evaluation, benchmarkPaper · Site
SpatialMedCTVQA, 3D spatial reasoning, benchmarkPaper

Benchmark Papers (44)

Safety, Robustness, and Fairness (34)

Themes Index

Cross-cutting topics that span multiple domain sections above. Each paper is listed once in its primary section; this index lets you find it by theme.

Fairness and Bias (7)

Hallucination and Reliability (39)

Safety and Robustness (31)

Privacy and Federated Learning (6)

RAG and Retrieval (52)

Multi-Agent Collaboration (145)

Truncated — view the full README on GitHub.

agents
ai-agents
awesome
awesome-list
computer-vision
deep-learning
healthcare
medical-image-analysis
medical-imaging
radiology
self-evolving
self-evolving-agents
trustworth-ai

Contributors

Nanboy-Ronan

100 commits

gexinh

1 commits

Languages

Python

100.0%