Latest Advances on Agentic AI & AI Agents for Healthcare
1,256
170 commits
updated Sep 8, 2026
This repository is a curated list of research papers, projects, and resources related to the application of Agentic AI / AI agents for healthcare, including medical image analysis, EHR manipulation, counseling, drug discovery, patient dialogue, and healthcare administration. AI agents refer to artificial intelligence systems that can autonomously perform tasks, make decisions, and interact with their environment, often through the use of large language models (LLMs), multi-agent systems, and tool integrations.
We will try to keep this list updated. If you find any errors or any missing papers, please don't hesitate to open issues or pull requests.
📘 Read our survey paper here: A Comprehensive Survey of AI Agents in Healthcare
If you find our paper and repository helpful, please cite:
@article{xu2026comprehensive,
title={A comprehensive survey of AI Agents in Healthcare},
author={Xu, Gelei and Li, Xueyang and Chen, Yixiong and Duan, Yuying and Wu, Shuqing and Yu, Haoxinran and Chiu, Ching-Hao and Ni, Juntong and Tang, Ningzhi and Li, Toby Jia-Jun and others},
journal={Journal of Biomedical Informatics},
pages={105045},
year={2026},
publisher={Elsevier}
}
(Agents designed to process and reason over multiple data types like images, text, and structured data)
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| MIRA: Medical Image Reflection for Agentic Diagnosis | arXiv | 2026.08 | Paper | Not Available |
| Inference-Time Agentic Decision Rules Beat Longer Evolving Search for Multi-Image Medical Reasoning | CVPR Workshop | 2026.07 | Paper | Not Available |
| Understanding From Human Perspective: A Multi-agent System for Interactive Egocentric Medical Image Segmentation | arXiv | 2026.07 | Paper | GitHub |
| MedRLM: Recursive Multimodal Health Intelligence for Long-Context Clinical Reasoning, Sensor-Guided Screening, Evidence-Grounded Decision Support, and Community-to-Tertiary Referral Optimization | arXiv | 2026.06 | Paper | Not Available |
| XMedFusion: A Knowledge-Guided Multimodal Perception and Reasoning Framework for Autonomous Medical Systems | arXiv | 2026.06 | Paper | Not Available |
| ArogyaSutra: A Multi-Agent Framework for Multimodal Medical Reasoning in Indic Languages | IJCAI | 2026.06 | Paper | Not Available |
| Towards Conversational Medical AI with Eyes, Ears and a Voice | arXiv | 2026.05 | Paper | Not Available |
| VERITAS: A Multi-Agent Co-Scientist for Verifiable Image-Derived Hypothesis Testing | arXiv | 2026.04 | Paper | GitHub |
| Camyla: Scaling Autonomous Research in Medical Image Segmentation | arXiv | 2026.04 | Paper | Project |
| MedVR: Annotation-Free Medical Visual Reasoning via Agentic Reinforcement Learning | ICLR | 2026.04 | Paper | Not Available |
| MedOpenClaw: Auditable Medical Imaging Agents Reasoning over Uncurated Full Studies | arXiv | 2026.03 | Paper | GitHub Project |
| Cerebra: A Multidisciplinary AI Board for Multimodal Dementia Characterization and Risk Assessment | arXiv | 2026.03 | Paper | Not Available |
| Shifting Adaptation from Weight Space to Memory Space: A Memory-Augmented Agent for Medical Image Segmentation | arXiv | 2026.03 | Paper | Not Available |
| Evolving Medical Imaging Agents via Experience-driven Self-skill Discovery | arXiv | 2026.03 | Paper | Not Available |
| Towards a Medical AI Scientist | arXiv | 2026.03 | Paper | Project |
| Meissa: Multi-modal Medical Agentic Intelligence | arXiv | 2026.03 | Paper | GitHub |
| CARE: Towards Clinical Accountability in Multi-Modal Medical Reasoning | ICLR | 2026.03 | Paper | Project |
| 3DMedAgent: Unified Perception-to-Understanding for 3D Medical Analysis | arXiv | 2026.02 | Paper | Not Available |
| CoMMa: Contribution-Aware Medical Multi-Agents From A Game-Theoretic Perspective | arXiv | 2026.02 | Paper | Not Available |
| MedXIAOHE: A Comprehensive Recipe for Building Medical MLLMs | arXiv | 2026.02 | Paper | Not Available |
| Picking the Right Specialist: Attentive Neural Process-based Selection of Task-Specialized Models | arXiv | 2026.02 | Paper | Not Available |
| Human-Guided Agentic AI for Multimodal Clinical Prediction | ICHI | 2026.02 | Paper | Not Available |
| MedSAM-Agent: Empowering Interactive Medical Image Segmentation with Multi-turn Agentic RL | arXiv | 2026.02 | Paper | GitHub |
| IBISAgent: Reinforcing Pixel-Level Visual Reasoning in MLLMs | arXiv | 2026.01 | Paper | Not Available |
| MedEyes: Learning Dynamic Visual Focus for Medical Progressive Diagnosis | arXiv | 2025.11 | Paper | GitHub |
| MedSAM3: Delving into Segment Anything with Medical Concepts | arXiv | 2025.11 | Paper | GitHub |
| AURA: A Multi-modal Medical Agent for Understanding, Reasoning & Annotation | MICCAI workshop | 2025.07 | Paper | GitHub |
| MedAgent-Pro: Towards Evidence-based Multi-modal Medical Diagnosis via Reasoning Agentic Workflow | arXiv | 2025.03 | Paper | GitHub |
| M^3Builder: A Multi-Agent System for Automated Machine Learning in Medical Imaging | arXiv | 2025.02 | Paper | GitHub |
| MAM: Modular Multi-Agent Framework for Multi-Modal Medical Diagnosis via Role-Specialized Collaboration | ACL | 2025 | Paper | GitHub |
| MedAgentSim: Self-Evolving Multi-Agent Simulations for Realistic Clinical Interactions | MICCAI | 2025 | Paper | GitHub |
| MDAgents: An Adaptive Collaboration of LLMs for Medical Decision-Making | NeurIPS (Oral) | 2024 | Paper | GitHub |
| MMedAgent: Learning to Use Medical Tools with Multi-modal Agent | EMNLP Findings | 2024 | Paper | GitHub |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| Policy-Driven CT-Agent: Modeling Phase-Aware Diagnostic Control for Clinically Consistent CT Reasoning | arXiv | 2026.07 | Paper | Not Available |
| CogRad: A Cognitively-Inspired Multi-Agent Framework for Radiology Report Generation | arXiv | 2026.07 | Paper | Not Available |
| A multi-agent system for spine MRI report generation from multi-sequence imaging | arXiv | 2026.06 | Paper | Not Available |
| ABRA: Agent Benchmark for Radiology Applications | arXiv | 2026.05 | Paper | Not Available |
| DeepTumorVQA: A Hierarchical 3D CT Benchmark for Stage-Wise Evaluation of Medical VLMs and Tool-Augmented Agents | arXiv | 2026.05 | Paper | Not Available |
| GAZE: Grounded Agentic Zero-shot Evaluation with Viewer-Level Tools and Literature Retrieval on Rare Brain MRI | arXiv | 2026.05 | Paper | Not Available |
| Agentic Large Language Models for Training-Free Neuro-Radiological Image Analysis | arXiv | 2026.04 | Paper | Not Available |
| MARCH: Multi-Agent Radiology Clinical Hierarchy for CT Report Generation | ACL | 2026.04 | Paper | Not Available |
| RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography | arXiv | 2026.04 | Paper | Not Available |
| Evo-MedAgent: Beyond One-Shot Diagnosis with Agents That Remember, Reflect, and Improve | arXiv | 2026.04 | Paper | Not Available |
| XrayClaw: Cooperative-Competitive Multi-Agent Alignment for Trustworthy Chest X-ray Diagnosis | arXiv | 2026.04 | Paper | Not Available |
| EviAgent: Evidence-Driven Agent for Radiology Report Generation | arXiv | 2026.03 | Paper | Not Available |
| Agentic Automation of BT-RADS Scoring: End-to-End Multi-Agent System for Standardized Brain Tumor Follow-up Assessment | arXiv | 2026.03 | Paper | Not Available |
| DUCX: Decomposing Unfairness in Tool-Using Chest X-ray Agents | arXiv | 2026.03 | Paper | Not Available |
| Can Agents Distinguish Visually Hard-to-Separate Diseases in a Zero-Shot Setting? | arXiv | 2026.02 | Paper | GitHub |
| Which Tool Response Should I Trust? Tool-Expertise-Aware CXR Agent with Multimodal Agentic Learning | arXiv | 2026.02 | Paper | Not Available |
| Perfusion Imaging and Single Material Reconstruction in Polychromatic Photon Counting CT | arXiv | 2026.02 | Paper | GitHub |
| Route, Retrieve, Reflect, Repair: Self-Improving Agentic Framework for Visual Detection | arXiv | 2026.01 | Paper | GitHub |
| Explainable Agentic AI Framework for Acute Ischemic Stroke Imaging Decisions | arXiv | 2026.01 | Paper | Not Available |
| LungNoduleAgent: A Collaborative Multi-Agent System for Precision Diagnosis of Lung Nodules | AAAI | 2026.1 | Paper | GitHub |
| Bidirectional human-AI collaboration in brain tumour assessments improves both expert human and AI agent performance | arXiv | 2025.12 | Paper | Not Available |
| INFORM-CT: INtegrating LLMs and VLMs FOR Incidental Findings Management in Abdominal CT | arXiv | 2025.12 | Paper | Not Available |
| Radiologist Copilot: Agentic AI Assistant for Holistic Radiology Reporting with Quality Control | arXiv | 2025.12 | Paper | Not Available |
| A Multi-Agent System for Complex Reasoning in Radiology Visual Question Answering | arXiv | 2025.08 | Paper | Not Available |
| AT-CXR: Uncertainty-Aware Agentic Triage for Chest X-rays | arXiv | 2025.08 | Paper | GitHub |
| PASS: Probabilistic Agentic Supernet Sampling for Interpretable and Adaptive Chest X-Ray Reasoning | arXiv | 2025.08 | Paper | GitHub |
| RadFabric: Agentic AI System with Reasoning Capability for Radiology | arXiv | 2025.06 | Paper | Project |
| A Multimodal Multi-Agent Framework for Radiology Report Generation | arXiv | 2025.05 | Paper | Not Available |
| CT-Agent: A Multimodal-LLM Agent for 3D CT Radiology Question Answering | arXiv | 2025.05 | Paper | Not Available |
| MedRAX: Medical reasoning agent for chest x-ray | ICML | 2025.02 | Paper | GitHub |
| Vision-language model for report generation and outcome prediction in CT pulmonary angiogram | npj Digital Medicine | 2025 | Paper | GitHub |
| AgentMRI: A Vison Language Model-Powered AI System for Self-regulating MRI Reconstruction with Multiple Degradations | Journal of imaging informatics in medicine | 2025 | Paper | Not Available |
| Enhancing LLMs for Impression Generation in Radiology Reports through a Multi-Agent System | arXiv | 2024.12 | Paper | Not Available |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| Trust but Verify:Evidence-Linked Multi-Agent Clinical Information Extraction in Pathology | arXiv | 2026.07 | Paper | Not Available |
| Democratizing and accelerating AI-driven pathology research through agentic intelligence | arXiv | 2026.06 | Paper | Not Available |
| Prompt, Plan, Extract: Zero-Shot Agentic LLMs Workflows for Lung Pathology Extraction from Clinical Narratives | arXiv | 2026.06 | Paper | Not Available |
| A Multi-modal Agentic Co-pilot for Evidence Grounded Computational Pathology | arXiv | 2026.06 | Paper | Not Available |
| Computational Pathology in the Era of Emerging Foundation and Agentic AI -- International Expert Perspectives | arXiv | 2026.03 | Paper | Not Available |
| LAMMI-Pathology: A Tool-Centric Bottom-Up LVLM-Agent Framework for Molecularly Informed Medical Intelligence | arXiv | 2026.02 | Paper | Not Available |
| SurvAgent: Hierarchical CoT-Enhanced Case Banking and Dichotomy-Based Multi-Agent System for Multimodal Survival Prediction | arXiv | 2025.11 | Paper | Not Available |
| GMAT: Grounded Multi-Agent Clinical Description Generation for Text Encoder in Vision-Language MIL | arXiv | 2025.08 | Paper | Not Available |
| Patho-AgenticRAG: Towards Multimodal Agentic Retrieval-Augmented Generation for Pathology VLMs | arXiv | 2025.08 | Paper | GitHub |
| Evidence-based diagnostic reasoning with multi-agent copilot for human pathology | arXiv | 2025.06 | Paper | Not Available |
| CPathAgent: An Agent-based Foundation Model for Interpretable High-Resolution Pathology Image Analysis | NeurIPS | 2025.05 | Paper | Not Available |
| PathFinder: A Multi-Modal Multi-Agent System for Medical Diagnostic Decision-Making Applied to Histopathology | ICCV | 2025.02 | Paper | project GitHub |
| WSI-Agents: A Collaborative Multi-Agent System for Multi-Modal Whole Slide Image Analysis | MICCAI (Oral) | 2025 | Paper | GitHub |
| Path-RAG: Knowledge-Guided Key Region Retrieval for Open-ended Pathology Visual Question Answering | MLHS | 2025 | Paper | GitHub |
| Pathgen-1.6m: 1.6 million pathology image-text pairs generation through multi-agent collaboration | ICLR (Oral) | 2024 | Paper | GitHub |
| PathAsst: A Generative Foundation AI Assistant towards Artificial General Intelligence of Pathology | AAAI | 2024 | Paper | GitHub |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| Diagnosing as Cardiologists Do: ECG Agents with Doctor-Grounded Priors for Clinical Reasoning Across Diseases and Populations | arXiv | 2026.08 | Paper | Not Available |
| Cardiologent: Multi-Agent Clinical Decision Support for Patient-Level Arrhythmia Assessment, Urgency, and Management | arXiv | 2026.07 | Paper | Not Available |
| ECG Foundation Models and Medical LLMs for Agentic Cardiovascular Intelligence at the Edge: A Review and Outlook | arXiv | 2026.04 | Paper | Not Available |
| Multi-Agent Reasoning for Cardiovascular Imaging Phenotype Analysis | MICCAI | 2025.07 | Paper | GitHub |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| Auditable agentic AI for evidence-grounded thyroid ultrasound diagnosis and reporting | arXiv | 2026.08 | Paper | Not Available |
| Echo-α: Large Agentic Multimodal Reasoning Model for Ultrasound Interpretation | arXiv | 2026.04 | Paper | Not Available |
| Anatomical Prior-Driven Framework for Autonomous Robotic Cardiac Ultrasound Standard View Acquisition | ICRA | 2026.03 | Paper | Not Available |
| Intelligent Virtual Sonographer (IVS): Enhancing Physician-Robot-Patient Communication | arXiv | 2025.07 | Paper | GitHub |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| Autonomous Radiotherapy Treatment Planning Using DOLA: A Privacy-Preserving, LLM-Based Optimization Agent | arXiv | 2025.03 | Paper | Not Available |
| A feasibility study of automating radiotherapy planning with large language model agents | Physics in Medicine & Biology | 2025 | Paper | Not Available |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| DermAgent: A Self-Reflective Agentic System for Dermatological Image Analysis with Multi-Tool Reasoning and Traceable Decision-Making | MICCAI | 2026.05 | Paper | Not Available |
| Conversational AI in health: Design considerations from a Wizard-of-Oz dermatology case study with users, clinicians and a medical LLM | CHI 'EA | 2024 | Paper | Not Available |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| OPGAgent: An Agent for Auditable Dental Panoramic X-ray Interpretation | arXiv | 2026.03 | Paper | Not Available |
| DentEval: Fine-tuning-Free Expert-Aligned Assessment in Dental Education via LLM Agents | MICCAI | 2025 | Paper | GitHub |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| CPAgents: Agentic Composite Phenotype Generation for Cardiac Disease Association | MICCAI | 2026.06 | Paper | Not Available |
| DeepBD: A Grounded Agentic Workflow for Variant Prioritization and Diagnosis of Genetic Birth Defects | arXiv | 2026.06 | Paper | Not Available |
| Autonomous Agent-Orchestrated Digital Twins (AADT): State Synchronization in Rare Genetic Disorders | arXiv | 2026.03 | Paper | Not Available |
| ProtRLSearch: A Multi-Round Multimodal Protein Search Agent with LLMs Trained via RL | arXiv | 2026.03 | Paper | Not Available |
| Geneagent: self-verification language agent for gene-set analysis using domain databases | Nature Methods | 2025 | Paper | GitHub |
| CRISPR-GPT for agentic automation of gene-editing experiments | Nature BME | 2025 | Paper | GitHub |
| HEAL-KGGen: A Hierarchical Multi-Agent LLM Framework for Genetic Biomarker-Based Medical Diagnosis | biorxiv | 2025 | Paper | GitHub |
| AI-HOPE: An AI-Driven conversational agent for enhanced clinical and genomic data integration | Bioinformatics | 2024.12 | Paper | GitHub |
| dna-claude-analysis: AI-powered personal genome analysis agent using Claude | GitHub | 2025 | Not Available | GitHub |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| A Multi-Agent System for Autonomous, Fine-Tuning-Free Clinical Symptom Detection: Development and Validation Study | arXiv | 2026.07 | Paper | Not Available |
| Multi-Large Language Model Orchestrated Severity Assessment of Clinical Records (MOSAIC) | arXiv | 2026.07 | Paper | Not Available |
| Configurable Clinical Information Extraction with Agentic RAG: What Works, What Breaks, and Why | arXiv | 2026.06 | Paper | Not Available |
| COTCAgent: Preventive Consultation via Probabilistic Chain-of-Thought Completion | arXiv | 2026.05 | Paper | Not Available |
| Reinforcement Learning for Tool-Calling Agents in Fast Healthcare Interoperability Resources (FHIR) | arXiv | 2026.05 | Paper | Not Available |
| Generating synthetic electronic health record data using agent-based models to evaluate machine learning robustness under mass casualty incidents | CHIL | 2026.05 | Paper | Not Available |
| CuraView: A Multi-Agent Framework for Medical Hallucination Detection with GraphRAG-Enhanced Knowledge Verification | arXiv | 2026.05 | Paper | Not Available |
| PhysicianBench: Evaluating LLM Agents in Real-World EHR Environments | arXiv | 2026.05 | Paper | Not Available |
| Clinically Interpretable Sepsis Early Warning via LLM-Guided Simulation of Temporal Physiological Dynamics | arXiv | 2026.04 | Paper | Not Available |
| BLUEmed: Retrieval-Augmented Multi-Agent Debate for Clinical Error Detection | IEEE ICHI | 2026.04 | Paper | Not Available |
| Beyond the Individual: Virtualizing Multi-Disciplinary Reasoning for Clinical Intake via Collaborative Agents | ACL'26 Findings | 2026.04 | Paper | GitHub |
| Symphony for Medical Coding: A Next-Generation Agentic System for Scalable and Explainable Medical Coding | arXiv | 2026.03 | Paper | Not Available |
| Can LLM Agents Generate Real-World Evidence? Evaluating Observational Studies in Medical Databases | arXiv | 2026.03 | Paper | GitHub |
| From Physician Expertise to Clinical Agents: Preserving, Standardizing, and Scaling Physicians' Medical Expertise | arXiv | 2026.03 | Paper | Not Available |
| Empowering Locally Deployable Medical Agent via State Enhanced Logical Skills for FHIR-based Clinical Tasks | arXiv | 2026.03 | Paper | Not Available |
| When OpenClaw Meets Hospital: Toward an Agentic Operating System for Dynamic Clinical Workflows | arXiv | 2026.03 | Paper | Not Available |
| TRACE: Temporal Reasoning via Agentic Context Evolution for Streaming EHRs | arXiv | 2026.02 | Paper | Not Available |
| AgentEHR: Advancing Autonomous Clinical Decision-Making via Retrospective Summarization | arXiv | 2026.01 | Paper | Not Available |
| ExperienceWeaver: Optimizing Small-sample Experience Learning for Clinical Text Improvement | arXiv | 2026.02 | Paper | Not Available |
| Hybrid-Code: A Privacy-Preserving, Redundant Multi-Agent Framework for Reliable Local Clinical Coding | arXiv | 2025.12 | Paper | Not Available |
| HARMON-E: Hierarchical Agentic Reasoning for Multimodal Oncology Notes to Extract Structured Data | arXiv | 2025.12 | Paper | Not Available |
| ClinNoteAgents: An LLM Multi-Agent System for Predicting and Interpreting Heart Failure 30-Day Readmission from Clinical Notes | arXiv | 2025.12 | Paper | Not Available |
| MedDCR: Learning to Design Agentic Workflows for Medical Coding | arXiv | 2025.11 | Paper | Not Available |
| OEMA: Ontology-Enhanced Multi-Agent Collaboration Framework for Zero-Shot Clinical Named Entity Recognition | arXiv | 2025.11 | Paper | Not Available |
| Grounded by Experience: Generative Healthcare Prediction Augmented with Hierarchical Agentic Retrieval | arXiv | 2025.11 | Paper | Not Available |
| Traj-CoA: Patient Trajectory Modeling via Chain-of-Agents for Lung Cancer Risk Prediction | NeurIPS'25 Workshop | 2025.10 | Paper | Not Available |
| Automated Clinical Problem Detection from SOAP Notes using a Collaborative Multi-Agent LLM Architecture | arXiv | 2025.08 | Paper | Not Available |
| SNOW: Agent-Based Feature Generation from Clinical Notes for Outcome Prediction | arXiv | 2025.08 | Paper | Project |
| Trustworthy Agents for Electronic Health Records through Confidence Estimation | arXiv | 2025.8 | Paper | GitHub |
| Infherno: End-to-end agent-based FHIR resource synthesis from free-form clinical notes | arXiv | 2025.07 | Paper | GitHub |
| From EHRs to Patient Pathways: Scalable Modeling of Longitudinal Health Trajectories with LLMs | arXiv | 2025.6 | Paper | Not Available |
| CARE-AD: a multi-agent large language model framework for Alzheimer’s disease prediction | npj Digital Medicine | 2025 | Paper | GitHub |
| Colacare: Enhancing electronic health record modeling through large language model-driven multi-agent collaboration | arXiv | 2024.10 | Paper | [project] |
| EHRFlow: A Large Language Model-Driven Iterative Multi-Agent Electronic Health Record Data Analysis Workflow | KDD'24 Workshop | 2024.06 | Paper | GitHub |
| A multi-agent architecture for privacy-preserving natural language interaction with FHIR-based electronic health records | IEEE SoftCOM | 2024 | Paper | Not Available |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| Agent-Guided Relational Concept Discovery: Toward Interpretable Surgical Margin Assessment | MICCAI | 2026.07 | Paper | Not Available |
| CSAP-Assist: Instrument-Agent Dialogue Empowered Vision-Language Models for Collaborative Surgical Action Planning | MICCAI | 2025 | Paper | GitHub |
| Privacy-Preserving Operating Room Workflow Analysis using Digital Twins | arXiv | 2025.4 | Paper | Not Available |
Related free course: BioDockify Learn - AI in Healthcare: Diagnosis to Drug Discovery - 24 free AI-narrated video lessons covering explainable AI in clinical settings (SHAP, GradCAM, GEMEX), EHR modeling, wearables, and clinical-judgment training with automation-bias scenarios.
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| MedEasy: Designing AI Standardized Patients for Clinical Consultation Training | arXiv | 2026.06 | Paper | Not Available |
| Rethinking Patient Education as Multi-turn Multi-modal Interaction | arXiv | 2026.04 | Paper | Not Available |
| Persona-Based Requirements Engineering for Explainable Multi-Agent Educational Systems: A Scenario Simulator for Clinical Reasoning Training | CSTE | 2026.04 | Paper | Not Available |
| Dialogue to Question Generation for Evidence-based Medical Guideline Agent Development | ML4H | 2026.03 | Paper | Not Available |
| An Agentic AI Framework for Training General Practitioner Student Skills | arXiv | 2025.12 | Paper | Not Available |
| MedTutor-R1: Socratic Personalized Medical Teaching with Multi-Agent Simulation | arXiv | 2025.12 | Paper | GitHub |
| Exploring Community-Powered Conversational Agent for Health Knowledge Acquisition | arXiv | 2025.12 | Paper | Not Available |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| MARC v1: An Open-Source Multi-Agent Framework for Clinical AI Reasoning and Coordination | arXiv | 2026.08 | Paper | GitHub |
| Social Chain of Thought: A Multi-Agent Architecture Grounded in Medical Differential Diagnosis Methodology | arXiv | 2026.08 | Paper | Not Available |
| Harnessing Disagreement: Detecting Correlated Agreement Blindness in Multi-Agent Triage | arXiv | 2026.07 | Paper | Not Available |
| MedCalc-Pro: Solving Complex Medical Calculations with LLM Agents | arXiv | 2026.07 | Paper | Not Available |
| DEEPMED Search: An Open-Source Agentic Platform for Medical Deep Research with Introspective Verification | IJCAI | 2026.06 | Paper | Not Available |
| MedGuards: Multi-Agent System for Reliable Medical Error Detection and Correction | arXiv | 2026.06 | Paper | GitHub |
| Evo-RAD: Navigating Rare Retinal Disease Diagnosis via Self-Evolving Agentic Retrieval | MICCAI | 2026.06 | Paper | GitHub |
| Agentic AI-based Framework for Mitigating Premature Diagnostic Handoff and Silent Hallucination in Healthcare Applications | arXiv | 2026.06 | Paper | Not Available |
| Teaching agentic AI to learn expert reasoning for rare disease diagnosis | arXiv | 2026.06 | Paper | Not Available |
| Let LLMs Judge Each Other: Multi-Agent Peer-Reviewed Reasoning for Medical Question Answering | arXiv | 2026.06 | Paper | Not Available |
| Trust but Verify: Mitigating Medical Hallucinations via Post-Hoc Adversarial Auditing and Multi-Agent Feedback Loops | arXiv | 2026.06 | Paper | Not Available |
| MedLatentDx: Latent Multi-Agent Communication for Cross-Hospital Rare-Disease Diagnosis | arXiv | 2026.06 | Paper | Not Available |
| Experience Makes Skillful: Enabling Generalizable Medical Agent Reasoning via Self-Evolving Skill Memory | arXiv | 2026.06 | Paper | Not Available |
| Baichuan-M4: A Clinical-Grade Medical Agent System for Continuous Care | arXiv | 2026.06 | Paper | Not Available |
| D2MDT: Department-aware Multidisciplinary Team Consultation with Deliberation for Efficient Clinical Prediction | arXiv | 2026.06 | Paper | Not Available |
| MeDxAgent: Multi-Agent Consultation for Interactive Medical Diagnosis | arXiv | 2026.06 | Paper | Not Available |
| SEMA-RAG: A Self-Evolving Multi-Agent Retrieval-Augmented Generation Framework for Medical Reasoning | arXiv | 2026.05 | Paper | Not Available |
| MedExAgent: Training LLM Agents to Ask, Examine, and Diagnose in Noisy Clinical Environments | arXiv | 2026.05 | Paper | Not Available |
| Thinking Like a Clinician: A Cognitive AI Agent for Clinical Diagnosis via Panoramic Profiling and Adversarial Debate | arXiv | 2026.04 | Paper | Not Available |
| Neuro-Symbolic Resolution of Recommendation Conflicts in Multimorbidity Clinical Guidelines | AAAI Bridge | 2026.04 | Paper | Not Available |
| DeepER-Med: Advancing Deep Evidence-Based Research in Medicine Through Agentic AI | arXiv | 2026.04 | Paper | Not Available |
| QuarkMedSearch: A Long-Horizon Deep Search Agent for Exploring Medical Intelligence | arXiv | 2026.04 | Paper | Not Available |
| Dialectic-Med: Mitigating Diagnostic Hallucinations via Counterfactual Adversarial Multi-Agent Debate | ACL | 2026.04 | Paper | Not Available |
| Joint Optimization of Reasoning and Dual-Memory for Self-Learning Diagnostic Agent | arXiv | 2026.04 | Paper | Not Available |
| CARE: Privacy-Compliant Agentic Reasoning with Evidence Discordance | arXiv | 2026.04 | Paper | Not Available |
| Improving Clinical Diagnosis with Counterfactual Multi-Agent Reasoning | arXiv | 2026.03 | Paper | Not Available |
| MediHive: A Decentralized Agent Collective for Medical Reasoning | IEEE ICHI | 2026.03 | Paper | Not Available |
| ClinicalAgents: Multi-Agent Orchestration for Clinical Decision Making with Dual-Memory | arXiv | 2026.03 | Paper | Not Available |
| Multi-Agent Reasoning with Consistency Verification Improves Uncertainty Calibration in Medical MCQA | arXiv | 2026.03 | Paper | Not Available |
| CarePilot: A Multi-Agent Framework for Long-Horizon Computer Task Automation in Healthcare | CVPR Findings | 2026.03 | Paper | Not Available |
| Unified-MAS: Universally Generating Domain-Specific Nodes for Empowering Automatic Multi-Agent Systems | arXiv | 2026.03 | Paper | GitHub |
| TheraAgent: Multi-Agent Framework with Self-Evolving Memory for PET Theranostics | arXiv | 2026.03 | Paper | Not Available |
| OpenHospital: A Thing-in-itself Arena for Evolving and Benchmarking LLM-based Collective Intelligence | arXiv | 2026.03 | Paper | Not Available |
| MedScope: Incentivizing "Think with Videos" for Clinical Reasoning via Coarse-to-Fine Tool Calling | arXiv | 2026.02 | Paper | Not Available |
| ATPO: Adaptive Tree Policy Optimization for Multi-Turn Medical Dialogue | ICLR | 2026.03 | Paper | Not Available |
| MedCoRAG: Interpretable Hepatology Diagnosis via Hybrid Evidence Retrieval and Multispecialty Consensus | arXiv | 2026.03 | Paper | Not Available |
| MedCollab: Causal-Driven Multi-Agent Collaboration for Full-Cycle Clinical Diagnosis | arXiv | 2026.03 | Paper | Not Available |
| From Conflict to Consensus: Boosting Medical Reasoning via Multi-Round Agentic RAG | arXiv | 2026.03 | Paper | GitHub |
| TARSE: Test-Time Adaptation via Retrieval of Skills and Experience for Reasoning Agents | arXiv | 2026.03 | Paper | Not Available |
| A Multi-Agent Framework for Interpreting Multivariate Physiological Time Series | arXiv | 2026.03 | Paper | Not Available |
| Do Mixed-Vendor Multi-Agent LLMs Improve Clinical Diagnosis? | EACL Workshop | 2026.03 | Paper | Not Available |
| MedClarify: An Information-Seeking AI Agent for Medical Diagnosis | arXiv | 2026.02 | Paper | Not Available |
| MedCoG: Maximizing LLM Inference Density in Medical Reasoning via Meta-Cognitive Regulation | arXiv | 2026.02 | Paper | Not Available |
| Closing Reasoning Gaps in Clinical Agents with Differential Reasoning Learning | arXiv | 2026.02 | Paper | Not Available |
| A Multi-Agent Framework for Medical AI: Leveraging GPT, LLaMA, and DeepSeek R1 | arXiv | 2026.02 | Paper | Not Available |
| Pruning Minimal Reasoning Graphs for Efficient Retrieval-Augmented Generation | arXiv | 2026.02 | Paper | Not Available |
| RE-MCDF: Closed-Loop Multi-Expert LLM Reasoning for Knowledge-Grounded Clinical Diagnosis | arXiv | 2026.02 | Paper | Not Available |
| Agentic Reasoning for Large Language Models | arXiv | 2026.01 | Paper | GitHub |
| EvoClinician: A Self-Evolving Agent for Multi-Turn Medical Diagnosis | arXiv | 2026.01 | Paper | GitHub |
| Scaling Medical Reasoning Verification via Tool-Integrated Reinforcement Learning | arXiv | 2026.01 | Paper | Not Available |
| DEEPMED: Building a Medical DeepResearch Agent via Multi-hop Med-Search Data | arXiv | 2026.01 | Paper | Not Available |
| Multi-Aspect Knowledge-Enhanced Medical Vision-Language Pretraining with Multi-Agent Data Generation | arXiv | 2025.12 | Paper | Not Available |
| Incentivizing Tool-augmented Thinking with Images for Medical Image Analysis | arXiv | 2025.12 | Paper | Not Available |
| AdaSearch: Balancing Parametric Knowledge and Search in Large Language Models via Reinforcement Learning | arXiv | 2025.12 | Paper | Github |
| Multi-Agent Medical Decision Consensus Matrix System: An Intelligent Collaborative Framework for Oncology MDT Consultations | arXiv | 2025.12 | Paper | Not Available |
| Multi-Agent Intelligence for Multidisciplinary Decision-Making in Gastrointestinal Oncology | arXiv | 2025.12 | Paper | Not Available |
| DART: Leveraging Multi-Agent Disagreement for Tool Recruitment in Multimodal Reasoning | arXiv | 2025.12 | Paper | Github |
| MCP-AI: Protocol-Driven Intelligence Framework for Autonomous Reasoning in Healthcare | arXiv | 2025.12 | Paper | Not Available |
| Many-to-One Adversarial Consensus: Exposing Multi-Agent Collusion Risks in AI-Based Healthcare | arXiv | 2025.12 | Paper | Not Available |
| Thucy: An LLM-based Multi-Agent System for Claim Verification across Relational Databases | AAAI Workshop | 2025.12 | Paper | Not Available |
| UCAgents: Unidirectional Convergence for Visual Evidence Anchored Multi-Agent Medical Decision-Making | arXiv | 2025.12 | Paper | GitHub |
| KOM: A Multi-Agent Artificial Intelligence System for Precision Management of Knee Osteoarthritis (KOA) | arXiv | 2025.11 | Paper | Not Available |
| KRAL: Knowledge and Reasoning Augmented Learning for LLM-assisted Clinical Antimicrobial Therapy | arXiv | 2025.11 | Paper | Not Available |
| MedResearcher-R1: Expert-Level Medical Deep Researcher via A Knowledge-Informed Trajectory Synthesis Framework | arXiv | 2025.8 | Paper | GitHub |
| ConfAgents: A Conformal-Guided Multi-Agent Framework for Cost-Efficient Medical Diagnosis | arXiv | 2025.8 | Paper | GitHub |
| Tree-of-Reasoning: Towards Complex Medical Diagnosis via Multi-Agent Reasoning with Evidence Tree | arXiv | 2025.8 | Paper | GitHub |
| End-to-End Agentic RAG System Training for Traceable Diagnostic Reasoning | arXiv | 2025.8 | Paper | GitHub |
| A Multi-Agent Approach to Neurological Clinical Reasoning | arXiv | 2025.8 | Paper | Not Available |
| KERAP: A knowledge-enhanced reasoning approach for accurate zero-shot diagnosis prediction | arXiv | 2025.7 | Paper | GitHub |
| MMedAgent-RL: Optimizing Multi-Agent Collaboration for Multimodal Medical Reasoning | arXiv | 2025.06 | Paper | Not Available |
| An agentic system for rare disease diagnosis with traceable reasoning | arXiv | 2025.6 | Paper | [demo] |
| MedOrch: Medical Diagnosis with Tool-Augmented Reasoning Agents for Flexible Extensibility | arXiv | 2025.6 | Paper | Not Available |
| The Optimization Paradox in Clinical AI Multi-Agent Systems | arXiv | 2025.6 | Paper | GitHub |
| DoctorAgent-RL: A Multi-Agent Collaborative Reinforcement Learning System for Multi-Turn Clinical Dialogue | EMNLP | 2025.5 | Paper | GitHub |
| Silence is Not Consensus: Disrupting Agreement Bias in Multi-Agent LLMs via Catfish Agent for Clinical Decision Making | arXiv | 2025.5 | Paper | Not Available |
| MDTeamGPT: A Self-Evolving LLM-Based Multi-Agent Framework for Multi-Disciplinary Team Medical Consultation | EMNLP | 2025.3 | Paper | GitHub |
| The Application of MATEC (Multi-AI Agent Team Care) Framework in Sepsis Care | arXiv | 2025.3 | Paper | Not Available |
| Agentic Medical Knowledge Graphs Enhance Medical Question Answering: Bridging the Gap Between LLMs and Evolving Medical Knowledge | EMNLP Findings | 2025.2 | Paper | GitHub |
| A Layered Debating Multi-Agent System for Similar Disease Diagnosis | NAACL | 2025 | Paper | Not Available |
| KG4Diagnosis: A Hierarchical Multi-Agent LLM Framework with Knowledge Graph Enhancement | arXiv | 2024.12 | Paper | Not Available |
| Zodiac: A Cardiologist-Level LLM Framework for Multi-Agent Diagnostics | arXiv | 2024.10 | Paper | Not Available |
| MedAgents: Large Language Models as Collaborators for Zero-shot Medical Reasoning | ACL 2024 Findings | 2023.11 | Paper | GitHub |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| ODRA: Synthesizing Cognitive Behavioral Therapy Sessions with Structured Chain-Of-Thought and Dynamic Patient Resistance | arXiv | 2026.08 | Paper | Not Available |
| Knowledge-augmented Agentic AI for Mental Health Medication Information Seeking | arXiv | 2026.06 | Paper | Not Available |
| A Multi-Agent Audit Framework for High-Stakes Reasoning: Evaluation and Interpretability in Clinical Mental Health Screening | arXiv | 2026.06 | Paper | Not Available |
| An Agentic LLM-Based Framework for Population-Scale Mental Health Screening | IEEE BigData | 2026.05 | Paper | Not Available |
| AI-Care: A Conversational Agentic System for Task Coordination in Alzheimer's Disease Care | arXiv | 2026.05 | Paper | Not Available |
| Design and Evaluation of a Culturally Adapted Multimodal Virtual Agent for PTSD Screening | arXiv | 2026.04 | Paper | Not Available |
| OMIND: Framework for Knowledge Grounded Finetuning and Multi-Turn Dialogue Benchmark for Mental Health LLMs | arXiv | 2026.03 | Paper | Not Available |
| YAQIN: Culturally Sensitive, Agentic AI for Mental Healthcare Support Among Muslim Women in the UK | arXiv | 2026.03 | Paper | Not Available |
| MIND: Unified Inquiry and Diagnosis RL for Psychiatric Consultation | arXiv | 2026.03 | Paper | Not Available |
| SynthAgent: A Multi-Agent LLM Framework for Realistic Patient Simulation | AAAI Workshop | 2026.02 | Paper | Not Available |
| Advancing AI Trustworthiness Through Patient Simulation for Antidepressant Selection | arXiv | 2026.02 | Paper | Not Available |
| DemMA: Dementia Multi-Turn Dialogue Agent with Expert-Guided Reasoning and Action Simulation | arXiv | 2026.01 | Paper | Not Available |
| CittaVerse (一念万相) | AI-powered reminiscence therapy platform for dementia/MCI using narrative identity, autobiographical memory scaffolding, and 6-dimension narrative quality scoring | arXiv (in prep) | Paper | GitHub |
| coTherapist: A Behavior-Aligned Small Language Model to Support Mental Healthcare Experts | arXiv | 2026.01 | Paper | Not Available |
| Towards Efficient and Robust Linguistic Emotion Diagnosis for Mental Health | arXiv | 2026.01 | Paper | Not Available |
| ChatThero: An LLM-Supported Chatbot for Behavior Change and Therapeutic Support in Addiction Recovery | arXiv | 2025.08 | Paper | GitHub Reproduce |
| VChatter: Exploring Generative Conversational Agents for Simulating Exposure Therapy to Reduce Social Anxiety | arXiv | 2025.06 | Paper | Not Available |
| AnnaAgent: Dynamic Evolution Agent System with Multi-Session Memory for Realistic Seeker Simulation | ACL Findings | 2025.06 | Paper | GitHub |
| MIND: Towards Immersive Psychological Healing with Multi-Agent Inner Dialogue | EMNLP Findings | 2025.02 | Paper | GitHub Reproduce |
| Cami: A counselor agent supporting motivational interviewing through state inference and topic exploration | ACL | 2025.02 | Paper | GitHub |
| Autocbt: An autonomous multi-agent framework for cognitive behavioral therapy in psychological counseling | arXiv | 2025.01 | Paper | Not Available |
| PsyDraw: A Multi-Agent Multimodal System for Mental Health Screening in Left-Behind Children | arXiv | 2024.12 | Paper | GitHub |
| Cactus: Towards psychological counseling conversations using cognitive behavioral theory | EMNLP Findings | 2024.07 | Paper | GitHub |
| MentalAgora: A Gateway to Advanced Personalized Care in Mental Health through Multi-Agent Debating | arXiv | 2024.07 | Paper | GitHub |
| Compeer: A generative conversational agent for proactive peer support | arXiv | 2024.07 | Paper | GitHub |
| Understanding the benefits and challenges of using large language model-based conversational agents for mental well-being support | AMIA Annual Symposium Proceedings | 2023.07 | Paper | Not Available |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| Towards Expert-level Medical AI for Real-time Video Consultations | arXiv | 2026.08 | Paper | Not Available |
| Guideline-as-Oracle: Zero-Annotation Training of an Ophthalmic Telephone Triage Agent | arXiv | 2026.08 | Paper | Not Available |
| Towards Conversational Medical AI with Eyes, Ears and a Voice | arXiv | 2026.05 | Paper | Not Available |
| SymptomAI: Toward a Conversational AI Agent for Everyday Symptom Assessment | arXiv | 2026.05 | Paper | Not Available |
| ClinicBot: A Guideline-Grounded Clinical Chatbot with Prioritized Evidence RAG and Verifiable Citations | arXiv | 2026.05 | Paper | Not Available |
| Statistics, Not Scale: Modular Medical Dialogue with Bayesian Belief Engine | arXiv | 2026.04 | Paper | Not Available |
| EMSDialog: Synthetic Multi-person Emergency Medical Service Dialogue Generation from Electronic Patient Care Reports via Multi-LLM Agents | ACL Findings | 2026.04 | Paper | Not Available |
| AI Agents for Conversational Patient Triage: Preliminary Simulation-Based Evaluation with Real-World EHR Data | arXiv | 2025.6 | paper | Not Available |
| A two-stage proactive dialogue generator for efficient clinical information collection | Expert Systems with Applications | 2025 | Paper | Not Available |
| PIORS: Personalized Intelligent Outpatient Reception based on Large Language Model with Multi-Agents Medical Scenario Simulation | ACL Findings | 2024.11 | Paper | GitHub |
| A language model--powered simulated patient with automated feedback for history taking: Prospective study | JMIR | 2024 | Paper | Not Available |
| Conversational health agents: a personalized large language model-powered agent framework | JAMIA Open | 2024 | Paper | GitHub |
| Talk2Care: Facilitating asynchronous patient-provider communication with large-language-model | arXiv | 2023.9 | Paper | Not Available |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| Traj-Evolve: A Self-Evolving Multi-Agent System for Patient Trajectory Modeling in Lung Cancer Early Detection | arXiv | 2026.06 | Paper | Not Available |
| Detecting Clinical Discrepancies in Health Coaching Agents: A Dual-Stream Memory and Reconciliation Architecture | arXiv | 2026.04 | Paper | Not Available |
| Agentic AI for Personalized Physiotherapy: A Multi-Agent Framework for Generative Video Training and Real-Time Pose Correction | ICDH IEEE | 2026.04 | Paper | Not Available |
| Sense Less, Infer More: Agentic Multimodal Transformers for Edge Medical Intelligence | arXiv | 2026.04 | Paper | Not Available |
| NutriOrion: A Hierarchical Multi-Agent Framework for Personalized Nutrition Intervention | arXiv | 2026.02 | Paper | Not Available |
| FinAgent: An Agentic AI Framework Integrating Personal Finance and Nutrition Planning | arXiv | 2025.12 | Paper | Not Available |
| On-device Large Multi-modal Agent for Human Activity Recognition | arXiv | 2025.12 | Paper | Not Available |
| Causal Reinforcement Learning based Agent-Patient Interaction with Clinical Domain Knowledge | arXiv | 2025.12 | Paper | Not Available |
| AI-VaxGuide: An Agentic RAG-Based LLM for Vaccination Decisions | arXiv | 2025.07 | Paper | huggingface |
| A Conversational Agent for Early Detection of Neurotoxic Effects of Medications through Automated Intensive Observation | PACIFIC SYMPOSIUM ON BIOCOMPUTING | 2024 | Paper | Not Available |
| An agentic AI tool for personalized body composition screening and metabolic health diagnostics | Not Available | 2026 | Not Available | Link |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| The Anatomy of a Personal Health Agent | arXiv | 2025.08 | Paper | Not Available |
| A general-purpose AI avatar in healthcare | arXiv | 2024.01 | Paper | Not Available |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| CASCADE: An Agentic Regulatory Network Framework for Patient-Data-Validated Downstream Perturbation Prediction | arXiv | 2026.08 | Paper | Not Available |
| An AI agent for treatment reasoning over a biomedical tool universe | arXiv | 2026.06 | Paper | GitHub |
| BioInsight: Multi-Agent Orchestration for Interactive Biomedical Knowledge Discovery | arXiv | 2026.06 | Paper | Not Available |
| DeepRoot: A KG-Coordinated Multi-Agent System for Therapeutic Reasoning over Historical Medical Texts | arXiv | 2026.06 | Paper | Not Available |
| Beyond Prompt-Based Planning: MCP-Native Graph Planning-based Biomedical Agent System | arXiv | 2026.06 | Paper | Not Available |
| A Versatile AI Agent for Rare Disease Diagnosis and Risk Gene Prioritization (Hygieia) | arXiv | 2026.05 | Paper | Not Available |
| FastOMOP: A Foundational Architecture for Reliable Agentic Real-World Evidence Generation on OMOP CDM data | arXiv | 2026.04 | Paper | Not Available |
| Vibe Medicine: Redefining Biomedical Research Through Human-AI Co-Work | arXiv | 2026.04 | Paper | Not Available |
| RexDrug: Reliable Multi-Drug Combination Extraction through Reasoning-Enhanced LLMs | arXiv | 2026.03 | Paper | GitHub |
| ALPACA: A Reinforcement Learning Environment for Medication Repurposing in Alzheimer's Disease | arXiv | 2026.02 | Paper | Not Available |
| Causal-Enhanced AI Agents for Medical Research Screening | arXiv | 2026.01 | Paper | Not Available |
| MedAI: Evaluating TxAgent's Therapeutic Agentic Reasoning in the NeurIPS CURE-Bench Competition | arXiv | 2025.12 | Paper | Benchmark & Competition |
| ToolUniverse: An open platform for democratizing AI scientists | arXiv | 2025.09 | Paper | GitHub |
| BioScientistAgent: Designing LLM-Biomedical Agents with KG-Augmented RL Reasoning Modules | biorxiv | 2025.08 | Paper | Not Available |
| RAG-Enhanced Collaborative LLM Agents for Drug Discovery | arXiv | 2025.02 | Paper | Not Available |
| Large Language Model Agent for Modular Task Execution in Drug Discovery | arXiv | 2025.07 | Paper | GitHub |
| AUTOCT: Automating Interpretable Clinical Trial Prediction with LLM Agents | arXiv | EMNLP | Paper | GitHub |
| Llm agent swarm for hypothesis-driven drug discovery | arXiv | 2025.04 | Paper | Not Available |
| Txgemma: Efficient and agentic llms for therapeutics | arXiv | 2025.04 | Paper | Not Available |
| TrialGenie: Empowering Clinical Trial Design with Agentic Intelligence and Real World Data | medRxiv | 2025.04 | Paper | Not Available |
| TxAgent: An AI agent for therapeutic reasoning across a universe of tools | arXiv | 2025.03 | Paper | GitHub |
| Drugagent: Automating ai-aided drug discovery programming through llm multi-agent collaboration | AAAI 2025 workshop AI4Research | 2024.11 | Paper | GitHub |
| Drugagent: Multi-agent large language model-based reasoning for drug-target interaction prediction | arXiv | 2024.08 | Paper | GitHub |
| PRISM: Patient Records Interpretation for Semantic Clinical Trial Matching using Large Language Models | npj Digital Medicine | 2024.01 | Paper | Not Available |
| MALADE: Orchestration of LLM-powered Agents with Retrieval Augmented Generation for Pharmacovigilance | MLHC | 2024 | Paper | GitHub |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| From Single Chatbots to Governed Agent Ecosystems: An Agentic AI Pattern Catalogue and Orchestration Framework for Mission-Critical Hospital Information Management Systems | arXiv | 2026.08 | Paper | Not Available |
| From Siloed Algorithms to Compliance-First Agentic Platforms: A Multi-Layered Architecture for Hospital AI Systems | arXiv | 2026.08 | Paper | Not Available |
| Toward Trustworthy Large Language Model Agents in Healthcare | arXiv | 2026.07 | Paper | GitHub |
| Why Trust Your Agent? Empirical Security Gains from TRiSM-Guided Agentic Workflows in Healthcare | arXiv |
Truncated — view the full README on GitHub.
Latest Advances on Agentic AI & AI Agents for Healthcare
1,256
170 commits
updated Sep 8, 2026
This repository is a curated list of research papers, projects, and resources related to the application of Agentic AI / AI agents for healthcare, including medical image analysis, EHR manipulation, counseling, drug discovery, patient dialogue, and healthcare administration. AI agents refer to artificial intelligence systems that can autonomously perform tasks, make decisions, and interact with their environment, often through the use of large language models (LLMs), multi-agent systems, and tool integrations.
We will try to keep this list updated. If you find any errors or any missing papers, please don't hesitate to open issues or pull requests.
📘 Read our survey paper here: A Comprehensive Survey of AI Agents in Healthcare
If you find our paper and repository helpful, please cite:
@article{xu2026comprehensive,
title={A comprehensive survey of AI Agents in Healthcare},
author={Xu, Gelei and Li, Xueyang and Chen, Yixiong and Duan, Yuying and Wu, Shuqing and Yu, Haoxinran and Chiu, Ching-Hao and Ni, Juntong and Tang, Ningzhi and Li, Toby Jia-Jun and others},
journal={Journal of Biomedical Informatics},
pages={105045},
year={2026},
publisher={Elsevier}
}
(Agents designed to process and reason over multiple data types like images, text, and structured data)
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| MIRA: Medical Image Reflection for Agentic Diagnosis | arXiv | 2026.08 | Paper | Not Available |
| Inference-Time Agentic Decision Rules Beat Longer Evolving Search for Multi-Image Medical Reasoning | CVPR Workshop | 2026.07 | Paper | Not Available |
| Understanding From Human Perspective: A Multi-agent System for Interactive Egocentric Medical Image Segmentation | arXiv | 2026.07 | Paper | GitHub |
| MedRLM: Recursive Multimodal Health Intelligence for Long-Context Clinical Reasoning, Sensor-Guided Screening, Evidence-Grounded Decision Support, and Community-to-Tertiary Referral Optimization | arXiv | 2026.06 | Paper | Not Available |
| XMedFusion: A Knowledge-Guided Multimodal Perception and Reasoning Framework for Autonomous Medical Systems | arXiv | 2026.06 | Paper | Not Available |
| ArogyaSutra: A Multi-Agent Framework for Multimodal Medical Reasoning in Indic Languages | IJCAI | 2026.06 | Paper | Not Available |
| Towards Conversational Medical AI with Eyes, Ears and a Voice | arXiv | 2026.05 | Paper | Not Available |
| VERITAS: A Multi-Agent Co-Scientist for Verifiable Image-Derived Hypothesis Testing | arXiv | 2026.04 | Paper | GitHub |
| Camyla: Scaling Autonomous Research in Medical Image Segmentation | arXiv | 2026.04 | Paper | Project |
| MedVR: Annotation-Free Medical Visual Reasoning via Agentic Reinforcement Learning | ICLR | 2026.04 | Paper | Not Available |
| MedOpenClaw: Auditable Medical Imaging Agents Reasoning over Uncurated Full Studies | arXiv | 2026.03 | Paper | GitHub Project |
| Cerebra: A Multidisciplinary AI Board for Multimodal Dementia Characterization and Risk Assessment | arXiv | 2026.03 | Paper | Not Available |
| Shifting Adaptation from Weight Space to Memory Space: A Memory-Augmented Agent for Medical Image Segmentation | arXiv | 2026.03 | Paper | Not Available |
| Evolving Medical Imaging Agents via Experience-driven Self-skill Discovery | arXiv | 2026.03 | Paper | Not Available |
| Towards a Medical AI Scientist | arXiv | 2026.03 | Paper | Project |
| Meissa: Multi-modal Medical Agentic Intelligence | arXiv | 2026.03 | Paper | GitHub |
| CARE: Towards Clinical Accountability in Multi-Modal Medical Reasoning | ICLR | 2026.03 | Paper | Project |
| 3DMedAgent: Unified Perception-to-Understanding for 3D Medical Analysis | arXiv | 2026.02 | Paper | Not Available |
| CoMMa: Contribution-Aware Medical Multi-Agents From A Game-Theoretic Perspective | arXiv | 2026.02 | Paper | Not Available |
| MedXIAOHE: A Comprehensive Recipe for Building Medical MLLMs | arXiv | 2026.02 | Paper | Not Available |
| Picking the Right Specialist: Attentive Neural Process-based Selection of Task-Specialized Models | arXiv | 2026.02 | Paper | Not Available |
| Human-Guided Agentic AI for Multimodal Clinical Prediction | ICHI | 2026.02 | Paper | Not Available |
| MedSAM-Agent: Empowering Interactive Medical Image Segmentation with Multi-turn Agentic RL | arXiv | 2026.02 | Paper | GitHub |
| IBISAgent: Reinforcing Pixel-Level Visual Reasoning in MLLMs | arXiv | 2026.01 | Paper | Not Available |
| MedEyes: Learning Dynamic Visual Focus for Medical Progressive Diagnosis | arXiv | 2025.11 | Paper | GitHub |
| MedSAM3: Delving into Segment Anything with Medical Concepts | arXiv | 2025.11 | Paper | GitHub |
| AURA: A Multi-modal Medical Agent for Understanding, Reasoning & Annotation | MICCAI workshop | 2025.07 | Paper | GitHub |
| MedAgent-Pro: Towards Evidence-based Multi-modal Medical Diagnosis via Reasoning Agentic Workflow | arXiv | 2025.03 | Paper | GitHub |
| M^3Builder: A Multi-Agent System for Automated Machine Learning in Medical Imaging | arXiv | 2025.02 | Paper | GitHub |
| MAM: Modular Multi-Agent Framework for Multi-Modal Medical Diagnosis via Role-Specialized Collaboration | ACL | 2025 | Paper | GitHub |
| MedAgentSim: Self-Evolving Multi-Agent Simulations for Realistic Clinical Interactions | MICCAI | 2025 | Paper | GitHub |
| MDAgents: An Adaptive Collaboration of LLMs for Medical Decision-Making | NeurIPS (Oral) | 2024 | Paper | GitHub |
| MMedAgent: Learning to Use Medical Tools with Multi-modal Agent | EMNLP Findings | 2024 | Paper | GitHub |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| Policy-Driven CT-Agent: Modeling Phase-Aware Diagnostic Control for Clinically Consistent CT Reasoning | arXiv | 2026.07 | Paper | Not Available |
| CogRad: A Cognitively-Inspired Multi-Agent Framework for Radiology Report Generation | arXiv | 2026.07 | Paper | Not Available |
| A multi-agent system for spine MRI report generation from multi-sequence imaging | arXiv | 2026.06 | Paper | Not Available |
| ABRA: Agent Benchmark for Radiology Applications | arXiv | 2026.05 | Paper | Not Available |
| DeepTumorVQA: A Hierarchical 3D CT Benchmark for Stage-Wise Evaluation of Medical VLMs and Tool-Augmented Agents | arXiv | 2026.05 | Paper | Not Available |
| GAZE: Grounded Agentic Zero-shot Evaluation with Viewer-Level Tools and Literature Retrieval on Rare Brain MRI | arXiv | 2026.05 | Paper | Not Available |
| Agentic Large Language Models for Training-Free Neuro-Radiological Image Analysis | arXiv | 2026.04 | Paper | Not Available |
| MARCH: Multi-Agent Radiology Clinical Hierarchy for CT Report Generation | ACL | 2026.04 | Paper | Not Available |
| RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography | arXiv | 2026.04 | Paper | Not Available |
| Evo-MedAgent: Beyond One-Shot Diagnosis with Agents That Remember, Reflect, and Improve | arXiv | 2026.04 | Paper | Not Available |
| XrayClaw: Cooperative-Competitive Multi-Agent Alignment for Trustworthy Chest X-ray Diagnosis | arXiv | 2026.04 | Paper | Not Available |
| EviAgent: Evidence-Driven Agent for Radiology Report Generation | arXiv | 2026.03 | Paper | Not Available |
| Agentic Automation of BT-RADS Scoring: End-to-End Multi-Agent System for Standardized Brain Tumor Follow-up Assessment | arXiv | 2026.03 | Paper | Not Available |
| DUCX: Decomposing Unfairness in Tool-Using Chest X-ray Agents | arXiv | 2026.03 | Paper | Not Available |
| Can Agents Distinguish Visually Hard-to-Separate Diseases in a Zero-Shot Setting? | arXiv | 2026.02 | Paper | GitHub |
| Which Tool Response Should I Trust? Tool-Expertise-Aware CXR Agent with Multimodal Agentic Learning | arXiv | 2026.02 | Paper | Not Available |
| Perfusion Imaging and Single Material Reconstruction in Polychromatic Photon Counting CT | arXiv | 2026.02 | Paper | GitHub |
| Route, Retrieve, Reflect, Repair: Self-Improving Agentic Framework for Visual Detection | arXiv | 2026.01 | Paper | GitHub |
| Explainable Agentic AI Framework for Acute Ischemic Stroke Imaging Decisions | arXiv | 2026.01 | Paper | Not Available |
| LungNoduleAgent: A Collaborative Multi-Agent System for Precision Diagnosis of Lung Nodules | AAAI | 2026.1 | Paper | GitHub |
| Bidirectional human-AI collaboration in brain tumour assessments improves both expert human and AI agent performance | arXiv | 2025.12 | Paper | Not Available |
| INFORM-CT: INtegrating LLMs and VLMs FOR Incidental Findings Management in Abdominal CT | arXiv | 2025.12 | Paper | Not Available |
| Radiologist Copilot: Agentic AI Assistant for Holistic Radiology Reporting with Quality Control | arXiv | 2025.12 | Paper | Not Available |
| A Multi-Agent System for Complex Reasoning in Radiology Visual Question Answering | arXiv | 2025.08 | Paper | Not Available |
| AT-CXR: Uncertainty-Aware Agentic Triage for Chest X-rays | arXiv | 2025.08 | Paper | GitHub |
| PASS: Probabilistic Agentic Supernet Sampling for Interpretable and Adaptive Chest X-Ray Reasoning | arXiv | 2025.08 | Paper | GitHub |
| RadFabric: Agentic AI System with Reasoning Capability for Radiology | arXiv | 2025.06 | Paper | Project |
| A Multimodal Multi-Agent Framework for Radiology Report Generation | arXiv | 2025.05 | Paper | Not Available |
| CT-Agent: A Multimodal-LLM Agent for 3D CT Radiology Question Answering | arXiv | 2025.05 | Paper | Not Available |
| MedRAX: Medical reasoning agent for chest x-ray | ICML | 2025.02 | Paper | GitHub |
| Vision-language model for report generation and outcome prediction in CT pulmonary angiogram | npj Digital Medicine | 2025 | Paper | GitHub |
| AgentMRI: A Vison Language Model-Powered AI System for Self-regulating MRI Reconstruction with Multiple Degradations | Journal of imaging informatics in medicine | 2025 | Paper | Not Available |
| Enhancing LLMs for Impression Generation in Radiology Reports through a Multi-Agent System | arXiv | 2024.12 | Paper | Not Available |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| Trust but Verify:Evidence-Linked Multi-Agent Clinical Information Extraction in Pathology | arXiv | 2026.07 | Paper | Not Available |
| Democratizing and accelerating AI-driven pathology research through agentic intelligence | arXiv | 2026.06 | Paper | Not Available |
| Prompt, Plan, Extract: Zero-Shot Agentic LLMs Workflows for Lung Pathology Extraction from Clinical Narratives | arXiv | 2026.06 | Paper | Not Available |
| A Multi-modal Agentic Co-pilot for Evidence Grounded Computational Pathology | arXiv | 2026.06 | Paper | Not Available |
| Computational Pathology in the Era of Emerging Foundation and Agentic AI -- International Expert Perspectives | arXiv | 2026.03 | Paper | Not Available |
| LAMMI-Pathology: A Tool-Centric Bottom-Up LVLM-Agent Framework for Molecularly Informed Medical Intelligence | arXiv | 2026.02 | Paper | Not Available |
| SurvAgent: Hierarchical CoT-Enhanced Case Banking and Dichotomy-Based Multi-Agent System for Multimodal Survival Prediction | arXiv | 2025.11 | Paper | Not Available |
| GMAT: Grounded Multi-Agent Clinical Description Generation for Text Encoder in Vision-Language MIL | arXiv | 2025.08 | Paper | Not Available |
| Patho-AgenticRAG: Towards Multimodal Agentic Retrieval-Augmented Generation for Pathology VLMs | arXiv | 2025.08 | Paper | GitHub |
| Evidence-based diagnostic reasoning with multi-agent copilot for human pathology | arXiv | 2025.06 | Paper | Not Available |
| CPathAgent: An Agent-based Foundation Model for Interpretable High-Resolution Pathology Image Analysis | NeurIPS | 2025.05 | Paper | Not Available |
| PathFinder: A Multi-Modal Multi-Agent System for Medical Diagnostic Decision-Making Applied to Histopathology | ICCV | 2025.02 | Paper | project GitHub |
| WSI-Agents: A Collaborative Multi-Agent System for Multi-Modal Whole Slide Image Analysis | MICCAI (Oral) | 2025 | Paper | GitHub |
| Path-RAG: Knowledge-Guided Key Region Retrieval for Open-ended Pathology Visual Question Answering | MLHS | 2025 | Paper | GitHub |
| Pathgen-1.6m: 1.6 million pathology image-text pairs generation through multi-agent collaboration | ICLR (Oral) | 2024 | Paper | GitHub |
| PathAsst: A Generative Foundation AI Assistant towards Artificial General Intelligence of Pathology | AAAI | 2024 | Paper | GitHub |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| Diagnosing as Cardiologists Do: ECG Agents with Doctor-Grounded Priors for Clinical Reasoning Across Diseases and Populations | arXiv | 2026.08 | Paper | Not Available |
| Cardiologent: Multi-Agent Clinical Decision Support for Patient-Level Arrhythmia Assessment, Urgency, and Management | arXiv | 2026.07 | Paper | Not Available |
| ECG Foundation Models and Medical LLMs for Agentic Cardiovascular Intelligence at the Edge: A Review and Outlook | arXiv | 2026.04 | Paper | Not Available |
| Multi-Agent Reasoning for Cardiovascular Imaging Phenotype Analysis | MICCAI | 2025.07 | Paper | GitHub |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| Auditable agentic AI for evidence-grounded thyroid ultrasound diagnosis and reporting | arXiv | 2026.08 | Paper | Not Available |
| Echo-α: Large Agentic Multimodal Reasoning Model for Ultrasound Interpretation | arXiv | 2026.04 | Paper | Not Available |
| Anatomical Prior-Driven Framework for Autonomous Robotic Cardiac Ultrasound Standard View Acquisition | ICRA | 2026.03 | Paper | Not Available |
| Intelligent Virtual Sonographer (IVS): Enhancing Physician-Robot-Patient Communication | arXiv | 2025.07 | Paper | GitHub |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| Autonomous Radiotherapy Treatment Planning Using DOLA: A Privacy-Preserving, LLM-Based Optimization Agent | arXiv | 2025.03 | Paper | Not Available |
| A feasibility study of automating radiotherapy planning with large language model agents | Physics in Medicine & Biology | 2025 | Paper | Not Available |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| DermAgent: A Self-Reflective Agentic System for Dermatological Image Analysis with Multi-Tool Reasoning and Traceable Decision-Making | MICCAI | 2026.05 | Paper | Not Available |
| Conversational AI in health: Design considerations from a Wizard-of-Oz dermatology case study with users, clinicians and a medical LLM | CHI 'EA | 2024 | Paper | Not Available |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| OPGAgent: An Agent for Auditable Dental Panoramic X-ray Interpretation | arXiv | 2026.03 | Paper | Not Available |
| DentEval: Fine-tuning-Free Expert-Aligned Assessment in Dental Education via LLM Agents | MICCAI | 2025 | Paper | GitHub |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| CPAgents: Agentic Composite Phenotype Generation for Cardiac Disease Association | MICCAI | 2026.06 | Paper | Not Available |
| DeepBD: A Grounded Agentic Workflow for Variant Prioritization and Diagnosis of Genetic Birth Defects | arXiv | 2026.06 | Paper | Not Available |
| Autonomous Agent-Orchestrated Digital Twins (AADT): State Synchronization in Rare Genetic Disorders | arXiv | 2026.03 | Paper | Not Available |
| ProtRLSearch: A Multi-Round Multimodal Protein Search Agent with LLMs Trained via RL | arXiv | 2026.03 | Paper | Not Available |
| Geneagent: self-verification language agent for gene-set analysis using domain databases | Nature Methods | 2025 | Paper | GitHub |
| CRISPR-GPT for agentic automation of gene-editing experiments | Nature BME | 2025 | Paper | GitHub |
| HEAL-KGGen: A Hierarchical Multi-Agent LLM Framework for Genetic Biomarker-Based Medical Diagnosis | biorxiv | 2025 | Paper | GitHub |
| AI-HOPE: An AI-Driven conversational agent for enhanced clinical and genomic data integration | Bioinformatics | 2024.12 | Paper | GitHub |
| dna-claude-analysis: AI-powered personal genome analysis agent using Claude | GitHub | 2025 | Not Available | GitHub |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| A Multi-Agent System for Autonomous, Fine-Tuning-Free Clinical Symptom Detection: Development and Validation Study | arXiv | 2026.07 | Paper | Not Available |
| Multi-Large Language Model Orchestrated Severity Assessment of Clinical Records (MOSAIC) | arXiv | 2026.07 | Paper | Not Available |
| Configurable Clinical Information Extraction with Agentic RAG: What Works, What Breaks, and Why | arXiv | 2026.06 | Paper | Not Available |
| COTCAgent: Preventive Consultation via Probabilistic Chain-of-Thought Completion | arXiv | 2026.05 | Paper | Not Available |
| Reinforcement Learning for Tool-Calling Agents in Fast Healthcare Interoperability Resources (FHIR) | arXiv | 2026.05 | Paper | Not Available |
| Generating synthetic electronic health record data using agent-based models to evaluate machine learning robustness under mass casualty incidents | CHIL | 2026.05 | Paper | Not Available |
| CuraView: A Multi-Agent Framework for Medical Hallucination Detection with GraphRAG-Enhanced Knowledge Verification | arXiv | 2026.05 | Paper | Not Available |
| PhysicianBench: Evaluating LLM Agents in Real-World EHR Environments | arXiv | 2026.05 | Paper | Not Available |
| Clinically Interpretable Sepsis Early Warning via LLM-Guided Simulation of Temporal Physiological Dynamics | arXiv | 2026.04 | Paper | Not Available |
| BLUEmed: Retrieval-Augmented Multi-Agent Debate for Clinical Error Detection | IEEE ICHI | 2026.04 | Paper | Not Available |
| Beyond the Individual: Virtualizing Multi-Disciplinary Reasoning for Clinical Intake via Collaborative Agents | ACL'26 Findings | 2026.04 | Paper | GitHub |
| Symphony for Medical Coding: A Next-Generation Agentic System for Scalable and Explainable Medical Coding | arXiv | 2026.03 | Paper | Not Available |
| Can LLM Agents Generate Real-World Evidence? Evaluating Observational Studies in Medical Databases | arXiv | 2026.03 | Paper | GitHub |
| From Physician Expertise to Clinical Agents: Preserving, Standardizing, and Scaling Physicians' Medical Expertise | arXiv | 2026.03 | Paper | Not Available |
| Empowering Locally Deployable Medical Agent via State Enhanced Logical Skills for FHIR-based Clinical Tasks | arXiv | 2026.03 | Paper | Not Available |
| When OpenClaw Meets Hospital: Toward an Agentic Operating System for Dynamic Clinical Workflows | arXiv | 2026.03 | Paper | Not Available |
| TRACE: Temporal Reasoning via Agentic Context Evolution for Streaming EHRs | arXiv | 2026.02 | Paper | Not Available |
| AgentEHR: Advancing Autonomous Clinical Decision-Making via Retrospective Summarization | arXiv | 2026.01 | Paper | Not Available |
| ExperienceWeaver: Optimizing Small-sample Experience Learning for Clinical Text Improvement | arXiv | 2026.02 | Paper | Not Available |
| Hybrid-Code: A Privacy-Preserving, Redundant Multi-Agent Framework for Reliable Local Clinical Coding | arXiv | 2025.12 | Paper | Not Available |
| HARMON-E: Hierarchical Agentic Reasoning for Multimodal Oncology Notes to Extract Structured Data | arXiv | 2025.12 | Paper | Not Available |
| ClinNoteAgents: An LLM Multi-Agent System for Predicting and Interpreting Heart Failure 30-Day Readmission from Clinical Notes | arXiv | 2025.12 | Paper | Not Available |
| MedDCR: Learning to Design Agentic Workflows for Medical Coding | arXiv | 2025.11 | Paper | Not Available |
| OEMA: Ontology-Enhanced Multi-Agent Collaboration Framework for Zero-Shot Clinical Named Entity Recognition | arXiv | 2025.11 | Paper | Not Available |
| Grounded by Experience: Generative Healthcare Prediction Augmented with Hierarchical Agentic Retrieval | arXiv | 2025.11 | Paper | Not Available |
| Traj-CoA: Patient Trajectory Modeling via Chain-of-Agents for Lung Cancer Risk Prediction | NeurIPS'25 Workshop | 2025.10 | Paper | Not Available |
| Automated Clinical Problem Detection from SOAP Notes using a Collaborative Multi-Agent LLM Architecture | arXiv | 2025.08 | Paper | Not Available |
| SNOW: Agent-Based Feature Generation from Clinical Notes for Outcome Prediction | arXiv | 2025.08 | Paper | Project |
| Trustworthy Agents for Electronic Health Records through Confidence Estimation | arXiv | 2025.8 | Paper | GitHub |
| Infherno: End-to-end agent-based FHIR resource synthesis from free-form clinical notes | arXiv | 2025.07 | Paper | GitHub |
| From EHRs to Patient Pathways: Scalable Modeling of Longitudinal Health Trajectories with LLMs | arXiv | 2025.6 | Paper | Not Available |
| CARE-AD: a multi-agent large language model framework for Alzheimer’s disease prediction | npj Digital Medicine | 2025 | Paper | GitHub |
| Colacare: Enhancing electronic health record modeling through large language model-driven multi-agent collaboration | arXiv | 2024.10 | Paper | [project] |
| EHRFlow: A Large Language Model-Driven Iterative Multi-Agent Electronic Health Record Data Analysis Workflow | KDD'24 Workshop | 2024.06 | Paper | GitHub |
| A multi-agent architecture for privacy-preserving natural language interaction with FHIR-based electronic health records | IEEE SoftCOM | 2024 | Paper | Not Available |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| Agent-Guided Relational Concept Discovery: Toward Interpretable Surgical Margin Assessment | MICCAI | 2026.07 | Paper | Not Available |
| CSAP-Assist: Instrument-Agent Dialogue Empowered Vision-Language Models for Collaborative Surgical Action Planning | MICCAI | 2025 | Paper | GitHub |
| Privacy-Preserving Operating Room Workflow Analysis using Digital Twins | arXiv | 2025.4 | Paper | Not Available |
Related free course: BioDockify Learn - AI in Healthcare: Diagnosis to Drug Discovery - 24 free AI-narrated video lessons covering explainable AI in clinical settings (SHAP, GradCAM, GEMEX), EHR modeling, wearables, and clinical-judgment training with automation-bias scenarios.
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| MedEasy: Designing AI Standardized Patients for Clinical Consultation Training | arXiv | 2026.06 | Paper | Not Available |
| Rethinking Patient Education as Multi-turn Multi-modal Interaction | arXiv | 2026.04 | Paper | Not Available |
| Persona-Based Requirements Engineering for Explainable Multi-Agent Educational Systems: A Scenario Simulator for Clinical Reasoning Training | CSTE | 2026.04 | Paper | Not Available |
| Dialogue to Question Generation for Evidence-based Medical Guideline Agent Development | ML4H | 2026.03 | Paper | Not Available |
| An Agentic AI Framework for Training General Practitioner Student Skills | arXiv | 2025.12 | Paper | Not Available |
| MedTutor-R1: Socratic Personalized Medical Teaching with Multi-Agent Simulation | arXiv | 2025.12 | Paper | GitHub |
| Exploring Community-Powered Conversational Agent for Health Knowledge Acquisition | arXiv | 2025.12 | Paper | Not Available |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| MARC v1: An Open-Source Multi-Agent Framework for Clinical AI Reasoning and Coordination | arXiv | 2026.08 | Paper | GitHub |
| Social Chain of Thought: A Multi-Agent Architecture Grounded in Medical Differential Diagnosis Methodology | arXiv | 2026.08 | Paper | Not Available |
| Harnessing Disagreement: Detecting Correlated Agreement Blindness in Multi-Agent Triage | arXiv | 2026.07 | Paper | Not Available |
| MedCalc-Pro: Solving Complex Medical Calculations with LLM Agents | arXiv | 2026.07 | Paper | Not Available |
| DEEPMED Search: An Open-Source Agentic Platform for Medical Deep Research with Introspective Verification | IJCAI | 2026.06 | Paper | Not Available |
| MedGuards: Multi-Agent System for Reliable Medical Error Detection and Correction | arXiv | 2026.06 | Paper | GitHub |
| Evo-RAD: Navigating Rare Retinal Disease Diagnosis via Self-Evolving Agentic Retrieval | MICCAI | 2026.06 | Paper | GitHub |
| Agentic AI-based Framework for Mitigating Premature Diagnostic Handoff and Silent Hallucination in Healthcare Applications | arXiv | 2026.06 | Paper | Not Available |
| Teaching agentic AI to learn expert reasoning for rare disease diagnosis | arXiv | 2026.06 | Paper | Not Available |
| Let LLMs Judge Each Other: Multi-Agent Peer-Reviewed Reasoning for Medical Question Answering | arXiv | 2026.06 | Paper | Not Available |
| Trust but Verify: Mitigating Medical Hallucinations via Post-Hoc Adversarial Auditing and Multi-Agent Feedback Loops | arXiv | 2026.06 | Paper | Not Available |
| MedLatentDx: Latent Multi-Agent Communication for Cross-Hospital Rare-Disease Diagnosis | arXiv | 2026.06 | Paper | Not Available |
| Experience Makes Skillful: Enabling Generalizable Medical Agent Reasoning via Self-Evolving Skill Memory | arXiv | 2026.06 | Paper | Not Available |
| Baichuan-M4: A Clinical-Grade Medical Agent System for Continuous Care | arXiv | 2026.06 | Paper | Not Available |
| D2MDT: Department-aware Multidisciplinary Team Consultation with Deliberation for Efficient Clinical Prediction | arXiv | 2026.06 | Paper | Not Available |
| MeDxAgent: Multi-Agent Consultation for Interactive Medical Diagnosis | arXiv | 2026.06 | Paper | Not Available |
| SEMA-RAG: A Self-Evolving Multi-Agent Retrieval-Augmented Generation Framework for Medical Reasoning | arXiv | 2026.05 | Paper | Not Available |
| MedExAgent: Training LLM Agents to Ask, Examine, and Diagnose in Noisy Clinical Environments | arXiv | 2026.05 | Paper | Not Available |
| Thinking Like a Clinician: A Cognitive AI Agent for Clinical Diagnosis via Panoramic Profiling and Adversarial Debate | arXiv | 2026.04 | Paper | Not Available |
| Neuro-Symbolic Resolution of Recommendation Conflicts in Multimorbidity Clinical Guidelines | AAAI Bridge | 2026.04 | Paper | Not Available |
| DeepER-Med: Advancing Deep Evidence-Based Research in Medicine Through Agentic AI | arXiv | 2026.04 | Paper | Not Available |
| QuarkMedSearch: A Long-Horizon Deep Search Agent for Exploring Medical Intelligence | arXiv | 2026.04 | Paper | Not Available |
| Dialectic-Med: Mitigating Diagnostic Hallucinations via Counterfactual Adversarial Multi-Agent Debate | ACL | 2026.04 | Paper | Not Available |
| Joint Optimization of Reasoning and Dual-Memory for Self-Learning Diagnostic Agent | arXiv | 2026.04 | Paper | Not Available |
| CARE: Privacy-Compliant Agentic Reasoning with Evidence Discordance | arXiv | 2026.04 | Paper | Not Available |
| Improving Clinical Diagnosis with Counterfactual Multi-Agent Reasoning | arXiv | 2026.03 | Paper | Not Available |
| MediHive: A Decentralized Agent Collective for Medical Reasoning | IEEE ICHI | 2026.03 | Paper | Not Available |
| ClinicalAgents: Multi-Agent Orchestration for Clinical Decision Making with Dual-Memory | arXiv | 2026.03 | Paper | Not Available |
| Multi-Agent Reasoning with Consistency Verification Improves Uncertainty Calibration in Medical MCQA | arXiv | 2026.03 | Paper | Not Available |
| CarePilot: A Multi-Agent Framework for Long-Horizon Computer Task Automation in Healthcare | CVPR Findings | 2026.03 | Paper | Not Available |
| Unified-MAS: Universally Generating Domain-Specific Nodes for Empowering Automatic Multi-Agent Systems | arXiv | 2026.03 | Paper | GitHub |
| TheraAgent: Multi-Agent Framework with Self-Evolving Memory for PET Theranostics | arXiv | 2026.03 | Paper | Not Available |
| OpenHospital: A Thing-in-itself Arena for Evolving and Benchmarking LLM-based Collective Intelligence | arXiv | 2026.03 | Paper | Not Available |
| MedScope: Incentivizing "Think with Videos" for Clinical Reasoning via Coarse-to-Fine Tool Calling | arXiv | 2026.02 | Paper | Not Available |
| ATPO: Adaptive Tree Policy Optimization for Multi-Turn Medical Dialogue | ICLR | 2026.03 | Paper | Not Available |
| MedCoRAG: Interpretable Hepatology Diagnosis via Hybrid Evidence Retrieval and Multispecialty Consensus | arXiv | 2026.03 | Paper | Not Available |
| MedCollab: Causal-Driven Multi-Agent Collaboration for Full-Cycle Clinical Diagnosis | arXiv | 2026.03 | Paper | Not Available |
| From Conflict to Consensus: Boosting Medical Reasoning via Multi-Round Agentic RAG | arXiv | 2026.03 | Paper | GitHub |
| TARSE: Test-Time Adaptation via Retrieval of Skills and Experience for Reasoning Agents | arXiv | 2026.03 | Paper | Not Available |
| A Multi-Agent Framework for Interpreting Multivariate Physiological Time Series | arXiv | 2026.03 | Paper | Not Available |
| Do Mixed-Vendor Multi-Agent LLMs Improve Clinical Diagnosis? | EACL Workshop | 2026.03 | Paper | Not Available |
| MedClarify: An Information-Seeking AI Agent for Medical Diagnosis | arXiv | 2026.02 | Paper | Not Available |
| MedCoG: Maximizing LLM Inference Density in Medical Reasoning via Meta-Cognitive Regulation | arXiv | 2026.02 | Paper | Not Available |
| Closing Reasoning Gaps in Clinical Agents with Differential Reasoning Learning | arXiv | 2026.02 | Paper | Not Available |
| A Multi-Agent Framework for Medical AI: Leveraging GPT, LLaMA, and DeepSeek R1 | arXiv | 2026.02 | Paper | Not Available |
| Pruning Minimal Reasoning Graphs for Efficient Retrieval-Augmented Generation | arXiv | 2026.02 | Paper | Not Available |
| RE-MCDF: Closed-Loop Multi-Expert LLM Reasoning for Knowledge-Grounded Clinical Diagnosis | arXiv | 2026.02 | Paper | Not Available |
| Agentic Reasoning for Large Language Models | arXiv | 2026.01 | Paper | GitHub |
| EvoClinician: A Self-Evolving Agent for Multi-Turn Medical Diagnosis | arXiv | 2026.01 | Paper | GitHub |
| Scaling Medical Reasoning Verification via Tool-Integrated Reinforcement Learning | arXiv | 2026.01 | Paper | Not Available |
| DEEPMED: Building a Medical DeepResearch Agent via Multi-hop Med-Search Data | arXiv | 2026.01 | Paper | Not Available |
| Multi-Aspect Knowledge-Enhanced Medical Vision-Language Pretraining with Multi-Agent Data Generation | arXiv | 2025.12 | Paper | Not Available |
| Incentivizing Tool-augmented Thinking with Images for Medical Image Analysis | arXiv | 2025.12 | Paper | Not Available |
| AdaSearch: Balancing Parametric Knowledge and Search in Large Language Models via Reinforcement Learning | arXiv | 2025.12 | Paper | Github |
| Multi-Agent Medical Decision Consensus Matrix System: An Intelligent Collaborative Framework for Oncology MDT Consultations | arXiv | 2025.12 | Paper | Not Available |
| Multi-Agent Intelligence for Multidisciplinary Decision-Making in Gastrointestinal Oncology | arXiv | 2025.12 | Paper | Not Available |
| DART: Leveraging Multi-Agent Disagreement for Tool Recruitment in Multimodal Reasoning | arXiv | 2025.12 | Paper | Github |
| MCP-AI: Protocol-Driven Intelligence Framework for Autonomous Reasoning in Healthcare | arXiv | 2025.12 | Paper | Not Available |
| Many-to-One Adversarial Consensus: Exposing Multi-Agent Collusion Risks in AI-Based Healthcare | arXiv | 2025.12 | Paper | Not Available |
| Thucy: An LLM-based Multi-Agent System for Claim Verification across Relational Databases | AAAI Workshop | 2025.12 | Paper | Not Available |
| UCAgents: Unidirectional Convergence for Visual Evidence Anchored Multi-Agent Medical Decision-Making | arXiv | 2025.12 | Paper | GitHub |
| KOM: A Multi-Agent Artificial Intelligence System for Precision Management of Knee Osteoarthritis (KOA) | arXiv | 2025.11 | Paper | Not Available |
| KRAL: Knowledge and Reasoning Augmented Learning for LLM-assisted Clinical Antimicrobial Therapy | arXiv | 2025.11 | Paper | Not Available |
| MedResearcher-R1: Expert-Level Medical Deep Researcher via A Knowledge-Informed Trajectory Synthesis Framework | arXiv | 2025.8 | Paper | GitHub |
| ConfAgents: A Conformal-Guided Multi-Agent Framework for Cost-Efficient Medical Diagnosis | arXiv | 2025.8 | Paper | GitHub |
| Tree-of-Reasoning: Towards Complex Medical Diagnosis via Multi-Agent Reasoning with Evidence Tree | arXiv | 2025.8 | Paper | GitHub |
| End-to-End Agentic RAG System Training for Traceable Diagnostic Reasoning | arXiv | 2025.8 | Paper | GitHub |
| A Multi-Agent Approach to Neurological Clinical Reasoning | arXiv | 2025.8 | Paper | Not Available |
| KERAP: A knowledge-enhanced reasoning approach for accurate zero-shot diagnosis prediction | arXiv | 2025.7 | Paper | GitHub |
| MMedAgent-RL: Optimizing Multi-Agent Collaboration for Multimodal Medical Reasoning | arXiv | 2025.06 | Paper | Not Available |
| An agentic system for rare disease diagnosis with traceable reasoning | arXiv | 2025.6 | Paper | [demo] |
| MedOrch: Medical Diagnosis with Tool-Augmented Reasoning Agents for Flexible Extensibility | arXiv | 2025.6 | Paper | Not Available |
| The Optimization Paradox in Clinical AI Multi-Agent Systems | arXiv | 2025.6 | Paper | GitHub |
| DoctorAgent-RL: A Multi-Agent Collaborative Reinforcement Learning System for Multi-Turn Clinical Dialogue | EMNLP | 2025.5 | Paper | GitHub |
| Silence is Not Consensus: Disrupting Agreement Bias in Multi-Agent LLMs via Catfish Agent for Clinical Decision Making | arXiv | 2025.5 | Paper | Not Available |
| MDTeamGPT: A Self-Evolving LLM-Based Multi-Agent Framework for Multi-Disciplinary Team Medical Consultation | EMNLP | 2025.3 | Paper | GitHub |
| The Application of MATEC (Multi-AI Agent Team Care) Framework in Sepsis Care | arXiv | 2025.3 | Paper | Not Available |
| Agentic Medical Knowledge Graphs Enhance Medical Question Answering: Bridging the Gap Between LLMs and Evolving Medical Knowledge | EMNLP Findings | 2025.2 | Paper | GitHub |
| A Layered Debating Multi-Agent System for Similar Disease Diagnosis | NAACL | 2025 | Paper | Not Available |
| KG4Diagnosis: A Hierarchical Multi-Agent LLM Framework with Knowledge Graph Enhancement | arXiv | 2024.12 | Paper | Not Available |
| Zodiac: A Cardiologist-Level LLM Framework for Multi-Agent Diagnostics | arXiv | 2024.10 | Paper | Not Available |
| MedAgents: Large Language Models as Collaborators for Zero-shot Medical Reasoning | ACL 2024 Findings | 2023.11 | Paper | GitHub |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| ODRA: Synthesizing Cognitive Behavioral Therapy Sessions with Structured Chain-Of-Thought and Dynamic Patient Resistance | arXiv | 2026.08 | Paper | Not Available |
| Knowledge-augmented Agentic AI for Mental Health Medication Information Seeking | arXiv | 2026.06 | Paper | Not Available |
| A Multi-Agent Audit Framework for High-Stakes Reasoning: Evaluation and Interpretability in Clinical Mental Health Screening | arXiv | 2026.06 | Paper | Not Available |
| An Agentic LLM-Based Framework for Population-Scale Mental Health Screening | IEEE BigData | 2026.05 | Paper | Not Available |
| AI-Care: A Conversational Agentic System for Task Coordination in Alzheimer's Disease Care | arXiv | 2026.05 | Paper | Not Available |
| Design and Evaluation of a Culturally Adapted Multimodal Virtual Agent for PTSD Screening | arXiv | 2026.04 | Paper | Not Available |
| OMIND: Framework for Knowledge Grounded Finetuning and Multi-Turn Dialogue Benchmark for Mental Health LLMs | arXiv | 2026.03 | Paper | Not Available |
| YAQIN: Culturally Sensitive, Agentic AI for Mental Healthcare Support Among Muslim Women in the UK | arXiv | 2026.03 | Paper | Not Available |
| MIND: Unified Inquiry and Diagnosis RL for Psychiatric Consultation | arXiv | 2026.03 | Paper | Not Available |
| SynthAgent: A Multi-Agent LLM Framework for Realistic Patient Simulation | AAAI Workshop | 2026.02 | Paper | Not Available |
| Advancing AI Trustworthiness Through Patient Simulation for Antidepressant Selection | arXiv | 2026.02 | Paper | Not Available |
| DemMA: Dementia Multi-Turn Dialogue Agent with Expert-Guided Reasoning and Action Simulation | arXiv | 2026.01 | Paper | Not Available |
| CittaVerse (一念万相) | AI-powered reminiscence therapy platform for dementia/MCI using narrative identity, autobiographical memory scaffolding, and 6-dimension narrative quality scoring | arXiv (in prep) | Paper | GitHub |
| coTherapist: A Behavior-Aligned Small Language Model to Support Mental Healthcare Experts | arXiv | 2026.01 | Paper | Not Available |
| Towards Efficient and Robust Linguistic Emotion Diagnosis for Mental Health | arXiv | 2026.01 | Paper | Not Available |
| ChatThero: An LLM-Supported Chatbot for Behavior Change and Therapeutic Support in Addiction Recovery | arXiv | 2025.08 | Paper | GitHub Reproduce |
| VChatter: Exploring Generative Conversational Agents for Simulating Exposure Therapy to Reduce Social Anxiety | arXiv | 2025.06 | Paper | Not Available |
| AnnaAgent: Dynamic Evolution Agent System with Multi-Session Memory for Realistic Seeker Simulation | ACL Findings | 2025.06 | Paper | GitHub |
| MIND: Towards Immersive Psychological Healing with Multi-Agent Inner Dialogue | EMNLP Findings | 2025.02 | Paper | GitHub Reproduce |
| Cami: A counselor agent supporting motivational interviewing through state inference and topic exploration | ACL | 2025.02 | Paper | GitHub |
| Autocbt: An autonomous multi-agent framework for cognitive behavioral therapy in psychological counseling | arXiv | 2025.01 | Paper | Not Available |
| PsyDraw: A Multi-Agent Multimodal System for Mental Health Screening in Left-Behind Children | arXiv | 2024.12 | Paper | GitHub |
| Cactus: Towards psychological counseling conversations using cognitive behavioral theory | EMNLP Findings | 2024.07 | Paper | GitHub |
| MentalAgora: A Gateway to Advanced Personalized Care in Mental Health through Multi-Agent Debating | arXiv | 2024.07 | Paper | GitHub |
| Compeer: A generative conversational agent for proactive peer support | arXiv | 2024.07 | Paper | GitHub |
| Understanding the benefits and challenges of using large language model-based conversational agents for mental well-being support | AMIA Annual Symposium Proceedings | 2023.07 | Paper | Not Available |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| Towards Expert-level Medical AI for Real-time Video Consultations | arXiv | 2026.08 | Paper | Not Available |
| Guideline-as-Oracle: Zero-Annotation Training of an Ophthalmic Telephone Triage Agent | arXiv | 2026.08 | Paper | Not Available |
| Towards Conversational Medical AI with Eyes, Ears and a Voice | arXiv | 2026.05 | Paper | Not Available |
| SymptomAI: Toward a Conversational AI Agent for Everyday Symptom Assessment | arXiv | 2026.05 | Paper | Not Available |
| ClinicBot: A Guideline-Grounded Clinical Chatbot with Prioritized Evidence RAG and Verifiable Citations | arXiv | 2026.05 | Paper | Not Available |
| Statistics, Not Scale: Modular Medical Dialogue with Bayesian Belief Engine | arXiv | 2026.04 | Paper | Not Available |
| EMSDialog: Synthetic Multi-person Emergency Medical Service Dialogue Generation from Electronic Patient Care Reports via Multi-LLM Agents | ACL Findings | 2026.04 | Paper | Not Available |
| AI Agents for Conversational Patient Triage: Preliminary Simulation-Based Evaluation with Real-World EHR Data | arXiv | 2025.6 | paper | Not Available |
| A two-stage proactive dialogue generator for efficient clinical information collection | Expert Systems with Applications | 2025 | Paper | Not Available |
| PIORS: Personalized Intelligent Outpatient Reception based on Large Language Model with Multi-Agents Medical Scenario Simulation | ACL Findings | 2024.11 | Paper | GitHub |
| A language model--powered simulated patient with automated feedback for history taking: Prospective study | JMIR | 2024 | Paper | Not Available |
| Conversational health agents: a personalized large language model-powered agent framework | JAMIA Open | 2024 | Paper | GitHub |
| Talk2Care: Facilitating asynchronous patient-provider communication with large-language-model | arXiv | 2023.9 | Paper | Not Available |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| Traj-Evolve: A Self-Evolving Multi-Agent System for Patient Trajectory Modeling in Lung Cancer Early Detection | arXiv | 2026.06 | Paper | Not Available |
| Detecting Clinical Discrepancies in Health Coaching Agents: A Dual-Stream Memory and Reconciliation Architecture | arXiv | 2026.04 | Paper | Not Available |
| Agentic AI for Personalized Physiotherapy: A Multi-Agent Framework for Generative Video Training and Real-Time Pose Correction | ICDH IEEE | 2026.04 | Paper | Not Available |
| Sense Less, Infer More: Agentic Multimodal Transformers for Edge Medical Intelligence | arXiv | 2026.04 | Paper | Not Available |
| NutriOrion: A Hierarchical Multi-Agent Framework for Personalized Nutrition Intervention | arXiv | 2026.02 | Paper | Not Available |
| FinAgent: An Agentic AI Framework Integrating Personal Finance and Nutrition Planning | arXiv | 2025.12 | Paper | Not Available |
| On-device Large Multi-modal Agent for Human Activity Recognition | arXiv | 2025.12 | Paper | Not Available |
| Causal Reinforcement Learning based Agent-Patient Interaction with Clinical Domain Knowledge | arXiv | 2025.12 | Paper | Not Available |
| AI-VaxGuide: An Agentic RAG-Based LLM for Vaccination Decisions | arXiv | 2025.07 | Paper | huggingface |
| A Conversational Agent for Early Detection of Neurotoxic Effects of Medications through Automated Intensive Observation | PACIFIC SYMPOSIUM ON BIOCOMPUTING | 2024 | Paper | Not Available |
| An agentic AI tool for personalized body composition screening and metabolic health diagnostics | Not Available | 2026 | Not Available | Link |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| The Anatomy of a Personal Health Agent | arXiv | 2025.08 | Paper | Not Available |
| A general-purpose AI avatar in healthcare | arXiv | 2024.01 | Paper | Not Available |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| CASCADE: An Agentic Regulatory Network Framework for Patient-Data-Validated Downstream Perturbation Prediction | arXiv | 2026.08 | Paper | Not Available |
| An AI agent for treatment reasoning over a biomedical tool universe | arXiv | 2026.06 | Paper | GitHub |
| BioInsight: Multi-Agent Orchestration for Interactive Biomedical Knowledge Discovery | arXiv | 2026.06 | Paper | Not Available |
| DeepRoot: A KG-Coordinated Multi-Agent System for Therapeutic Reasoning over Historical Medical Texts | arXiv | 2026.06 | Paper | Not Available |
| Beyond Prompt-Based Planning: MCP-Native Graph Planning-based Biomedical Agent System | arXiv | 2026.06 | Paper | Not Available |
| A Versatile AI Agent for Rare Disease Diagnosis and Risk Gene Prioritization (Hygieia) | arXiv | 2026.05 | Paper | Not Available |
| FastOMOP: A Foundational Architecture for Reliable Agentic Real-World Evidence Generation on OMOP CDM data | arXiv | 2026.04 | Paper | Not Available |
| Vibe Medicine: Redefining Biomedical Research Through Human-AI Co-Work | arXiv | 2026.04 | Paper | Not Available |
| RexDrug: Reliable Multi-Drug Combination Extraction through Reasoning-Enhanced LLMs | arXiv | 2026.03 | Paper | GitHub |
| ALPACA: A Reinforcement Learning Environment for Medication Repurposing in Alzheimer's Disease | arXiv | 2026.02 | Paper | Not Available |
| Causal-Enhanced AI Agents for Medical Research Screening | arXiv | 2026.01 | Paper | Not Available |
| MedAI: Evaluating TxAgent's Therapeutic Agentic Reasoning in the NeurIPS CURE-Bench Competition | arXiv | 2025.12 | Paper | Benchmark & Competition |
| ToolUniverse: An open platform for democratizing AI scientists | arXiv | 2025.09 | Paper | GitHub |
| BioScientistAgent: Designing LLM-Biomedical Agents with KG-Augmented RL Reasoning Modules | biorxiv | 2025.08 | Paper | Not Available |
| RAG-Enhanced Collaborative LLM Agents for Drug Discovery | arXiv | 2025.02 | Paper | Not Available |
| Large Language Model Agent for Modular Task Execution in Drug Discovery | arXiv | 2025.07 | Paper | GitHub |
| AUTOCT: Automating Interpretable Clinical Trial Prediction with LLM Agents | arXiv | EMNLP | Paper | GitHub |
| Llm agent swarm for hypothesis-driven drug discovery | arXiv | 2025.04 | Paper | Not Available |
| Txgemma: Efficient and agentic llms for therapeutics | arXiv | 2025.04 | Paper | Not Available |
| TrialGenie: Empowering Clinical Trial Design with Agentic Intelligence and Real World Data | medRxiv | 2025.04 | Paper | Not Available |
| TxAgent: An AI agent for therapeutic reasoning across a universe of tools | arXiv | 2025.03 | Paper | GitHub |
| Drugagent: Automating ai-aided drug discovery programming through llm multi-agent collaboration | AAAI 2025 workshop AI4Research | 2024.11 | Paper | GitHub |
| Drugagent: Multi-agent large language model-based reasoning for drug-target interaction prediction | arXiv | 2024.08 | Paper | GitHub |
| PRISM: Patient Records Interpretation for Semantic Clinical Trial Matching using Large Language Models | npj Digital Medicine | 2024.01 | Paper | Not Available |
| MALADE: Orchestration of LLM-powered Agents with Retrieval Augmented Generation for Pharmacovigilance | MLHC | 2024 | Paper | GitHub |
| Title | Venue | Date | Paper Link | Project Page |
|---|---|---|---|---|
| From Single Chatbots to Governed Agent Ecosystems: An Agentic AI Pattern Catalogue and Orchestration Framework for Mission-Critical Hospital Information Management Systems | arXiv | 2026.08 | Paper | Not Available |
| From Siloed Algorithms to Compliance-First Agentic Platforms: A Multi-Layered Architecture for Hospital AI Systems | arXiv | 2026.08 | Paper | Not Available |
| Toward Trustworthy Large Language Model Agents in Healthcare | arXiv | 2026.07 | Paper | GitHub |
| Why Trust Your Agent? Empirical Security Gains from TRiSM-Guided Agentic Workflows in Healthcare | arXiv |
Truncated — view the full README on GitHub.