Skip to content

AgenticHealthAI/Awesome-AI-Agents-for-Healthcare

Folders and files

NameName
Last commit message
Last commit date

Latest commit

 

History

75 Commits
 
 
 
 
 
 

Repository files navigation

Awesome AI Agents for Healthcare

Awesome PRs Welcome Read Our Survey Star on GitHub

This repository is a curated list of research papers, projects, and resources related to the application of Agentic AI / AI agents for healthcare, including medical image analysis, EHR manipulation, counseling, drug discovery, patient dialogue, and healthcare administration. AI agents refer to artificial intelligence systems that can autonomously perform tasks, make decisions, and interact with their environment, often through the use of large language models (LLMs), multi-agent systems, and tool integrations.

  1. Below image introduces a comprehensive conceptual framework. It provides a holistic view, detailing the pipeline from initial data perception and foundational agent capabilities to a hierarchical application ecosystem.

Overall Landscape

  1. We conducted a quantitative analysis of recent academic literature, with the key findings summarized in below image. This analysis provides a data-driven snapshot of the field’s growth trajectory, technological underpinnings, and application hotspots:
  • Top Data Modalities: Textual data remain the most frequently utilized modalities. Time-Series and Genomics exhibit a high proportion of publications from 2025.
  • Top Technologies: The technological focus is heavily concentrated on three topics: 1) developing highlevel Frameworks, 2) enhancing agent Reasoning, and 3) designing Multi-Agent collaboration paradigms.
  • Top Application Domains: Agentic AI continues to be widely applied in broad domains like General Medicine, Public Health, and Mental Health. Drug Discovery and Genomics as particularly new frontiers.

Statistics for Research Trends

We will try to keep this list updated. If you find any errors or any missing paper, please don't hesitate to open issues or pull requests.

📘 Read our survey paper here: A Comprehensive Survey of AI Agents for Healthcare

If you find our paper and repository helpful, please cite:

@article{xu2025comprehensive,
  title={A Comprehensive Survey of Agentic AI in Healthcare},
  author={Xu, Gelei and Li, Xueyang and Chen, Yixiong and Duan, Yuying and Wu, Shuqing and Yu, Alexander and Chiu, Ching-Hao and Ni, Juntong and Tang, Ningzhi and Li, Toby Jia-Jun and others},
  journal={Authorea Preprints},
  year={2025},
  publisher={Authorea}
}

Table of Contents


Latest Papers

Year 2026

  1. [AAAI 2026] LungNoduleAgent: A Collaborative Multi-Agent System for Precision Diagnosis of Lung Nodules [paper] [Github]

Year 2025

  1. [arxiv 2025.12] Hybrid-Code: A Privacy-Preserving, Redundant Multi-Agent Framework for Reliable Local Clinical Coding [paper]
  2. [arxiv 2025.12] ClinDEF: A Dynamic Evaluation Framework for Large Language Models in Clinical Reasoning [paper]
  3. [arxiv 2025.12] HARMON-E: Hierarchical Agentic Reasoning for Multimodal Oncology Notes to Extract Structured Data [paper]
  4. [arxiv 2025.12] Bidirectional human-AI collaboration in brain tumour assessments improves both expert human and AI agent performance [paper]
  5. [arxiv 2025.12] On-device Large Multi-modal Agent for Human Activity Recognition [paper]
  6. [arxiv 2025.12] Scalably Enhancing the Clinical Validity of a Task Benchmark with Physician Oversight [paper]
  7. [arxiv 2025.12] Agent-Based Output Drift Detection for Breast Cancer Response Prediction in a Multisite Clinical Decision Support System [paper]
  8. [arxiv 2025.12] An Agentic AI Framework for Training General Practitioner Student Skills [paper]
  9. [arxiv 2025.12] ReX-MLE: The Autonomous Agent Benchmark for Medical Imaging Challenges [paper]
  10. [arxiv 2025.12] AdaSearch: Balancing Parametric Knowledge and Search in Large Language Models via Reinforcement Learning [paper]
  11. [arxiv 2025.12] A Multi-Agent Large Language Model Framework for Automated Qualitative Analysis [paper]
  12. [arxiv 2025.12] Mapis: A Knowledge-Graph Grounded Multi-Agent Framework for Evidence-Based PCOS Diagnosis [paper]
  13. [arxiv 2025.12] INFORM-CT: INtegrating LLMs and VLMs FOR Incidental Findings Management in Abdominal CT [paper]
  14. [arxiv 2025.12] Multi-Agent Medical Decision Consensus Matrix System: An Intelligent Collaborative Framework for Oncology MDT Consultations [paper]
  15. [arxiv 2025.12] Incentivizing Tool-augmented Thinking with Images for Medical Image Analysis [paper]
  16. [arxiv 2025.12] MedInsightBench: Evaluating Medical Analytics Agents Through Multi-Step Insight Discovery in Multimodal Medical Data [paper]
  17. [arxiv 2025.12] Socratic Students: Teaching Language Models to Learn by Asking Questions [paper]
  18. [arxiv 2025.12] MedAI: Evaluating TxAgent's Therapeutic Agentic Reasoning in the NeurIPS CURE-Bench Competition [paper] [Benchmark & Competition]
  19. [arxiv 2025.12] CP-Env: Evaluating Large Language Models on Clinical Pathways in a Controllable Hospital Environment [paper] [Github]
  20. [arxiv 2025.12] AutoMedic: An Automated Evaluation Framework for Clinical Conversational Agents with Medical Dataset Grounding [paper]
  21. [arxiv 2025.12] Exploring Community-Powered Conversational Agent for Health Knowledge Acquisition: A Case Study in Colorectal Cancer [paper]
  22. [arxiv 2025.12] Multi-Agent Intelligence for Multidisciplinary Decision-Making in Gastrointestinal Oncology [paper]
  23. [arxiv 2025.12] DART: Leveraging Multi-Agent Disagreement for Tool Recruitment in Multimodal Reasoning [paper]
  24. [arxiv 2025.12] ClinNoteAgents: An LLM Multi-Agent System for Predicting and Interpreting Heart Failure 30-Day Readmission from Clinical Notes [paper]
  25. [arxiv 2025.12] MedTutor-R1: Socratic Personalized Medical Teaching with Multi-Agent Simulation [paper]
  26. [arxiv 2025.12] MCP-AI: Protocol-Driven Intelligence Framework for Autonomous Reasoning in Healthcare [paper]
  27. [ICCV 2025 Highlight] Multi-Aspect Knowledge-Enhanced Medical Vision-Language Pretraining with Multi-Agent Data Generation [paper] [Github]
  28. [arxiv 2025.12] Thucy: An LLM-based Multi-Agent System for Claim Verification across Relational Databases [paper]
  29. [arxiv 2025.12] Many-to-One Adversarial Consensus: Exposing Multi-Agent Collusion Risks in AI-Based Healthcare [paper]
  30. [arxiv 2025.12] FinAgent: An Agentic AI Framework Integrating Personal Finance and Nutrition Planning [paper]
  31. [arxiv 2025.12] Radiologist Copilot: Agentic AI Assistant for Holistic Radiology Reporting with Quality Control [paper]
  32. [arxiv 2025.12] UCAgents: Unidirectional Convergence for Visual Evidence Anchored Multi-Agent Medical Decision-Making [paper]
  33. [arxiv 2025.12] First, do NOHARM: towards clinically safe large language models [paper]
  34. [arxiv 2025.12] Causal Reinforcement Learning based Agent-Patient Interaction with Clinical Domain Knowledge [paper]
  35. [arxiv 2025.11] MedSAM3: Delving into Segment Anything with Medical Concepts [paper] [Github]
  36. [arxiv 2025.11] SurvAgent: Hierarchical CoT-Enhanced Case Banking and Dichotomy-Based Multi-Agent System for Multimodal Survival Prediction [paper]
  37. [arxiv 2025.11] KOM: A Multi-Agent Artificial Intelligence System for Precision Management of Knee Osteoarthritis (KOA) [paper]
  38. [arxiv 2025.11] KRAL: Knowledge and Reasoning Augmented Learning for LLM-assisted Clinical Antimicrobial Therapy [paper]
  39. [arxiv 2025.11] Medical Malice: A Dataset for Context-Aware Safety in Healthcare LLMs [paper]
  40. [arxiv 2025.11] MedBench v4: A Robust and Scalable Benchmark for Evaluating Chinese Medical Language Models, Multimodal Models, and Intelligent Agents [paper]
  41. [arxiv 2025.11] Fair-GNE: Generalized Nash Equilibrium-Seeking Fairness in Multiagent Healthcare Automation [paper]
  42. [arxiv 2025.11] MedDCR: Learning to Design Agentic Workflows for Medical Coding [paper]
  43. [arxiv 2025.11] Grounded by Experience: Generative Healthcare Prediction Augmented with Hierarchical Agentic Retrieval [paper]
  44. [arxiv 2025.11] OEMA: Ontology-Enhanced Multi-Agent Collaboration Framework for Zero-Shot Clinical Named Entity Recognition [paper]
  45. [arxiv 2025.11] MedBuild AI: An Agent-Based Hybrid Intelligence Framework for Reshaping Agency in Healthcare Infrastructure Planning through Generative Design for Medical Architecture [paper]
  46. [arxiv 2025.11] From Passive to Proactive: A Multi-Agent System with Dynamic Task Orchestration for Intelligent Medical Pre-Consultation [paper]
  47. [arxiv 2025.11] Fine-Tuning DialoGPT on Common Diseases in Rural Nepal for Medical Conversations [paper]
  48. [arxiv 2025.10] Traj-CoA: Patient Trajectory Modeling via Chain-of-Agents for Lung Cancer Risk Prediction [paper]
  49. [arxiv 2025.10] FT-ARM: Fine-Tuned Agentic Reflection Multimodal Language Model for Pressure Ulcer Severity Classification with Reasoning [paper]
  50. [arxiv 2025.10] SNOMED CT-powered Knowledge Graphs for Structured Clinical Data and Diagnostic Reasoning [paper]
  51. [arxiv 2025.10] Speculative Model Risk in Healthcare AI: Using Storytelling to Surface Unintended Harms [paper]
  52. [arxiv 2025.10] MedCoAct: Confidence-Aware Multi-Agent Collaboration for Complete Clinical Decision [paper]
  53. [arxiv 2025.10] Haibu Mathematical-Medical Intelligent Agent:Enhancing Large Language Model Reliability in Medical Tasks via Verifiable Reasoning Chains [paper]
  54. [arxiv 2025.10] Reinforcement Learning for Clinical Reasoning: Aligning LLMs with ACR Imaging Appropriateness Criteria [paper]
  55. [arxiv 2025.10] CLARITY: Clinical Assistant for Routing, Inference, and Triage [paper]
  56. [arxiv 2025.10] Secure Multi-Modal Data Fusion in Federated Digital Health Systems via MCP [paper]
  57. [arxiv 2025.9] AgenticAD: A Specialized Multiagent System Framework for Holistic Alzheimer Disease Management [paper]
  58. [arxiv 2025.9] Agentic-AI Healthcare: Multilingual, Privacy-First Framework with MCP Agents [paper]
  59. [arxiv 2025.9] Online Decision Making with Generative Action Sets [paper]
  60. [arxiv 2025.9] PAME-AI: Patient Messaging Creation and Optimization using Agentic AI [paper]
  61. [arxiv 2025.9] A co-evolving agentic AI system for medical imaging analysis [paper]
  62. [arxiv 2025.9] FHIR-AgentBench: Benchmarking LLM Agents for Realistic Interoperable EHR Question Answering [paper] [Github]
  63. [arxiv 2025.9] MedFact: Benchmarking the Fact-Checking Capabilities of Large Language Models on Chinese Medical Texts [paper]
  64. [arxiv 2025.9] Agentic Temporal Graph of Reasoning with Multimodal Language Models: A Potential AI Aid to Healthcare [paper]
  65. [arxiv 2025.9] Using AI to Optimize Patient Transfer and Resource Utilization During Mass-Casualty Incidents: A Simulation Platform [paper]
  66. [arxiv 2025.9] Demo: Healthcare Agent Orchestrator (HAO) for Patient Summarization in Molecular Tumor Boards [paper]
  67. [arxiv 2025.9] Chatbot To Help Patients Understand Their Health [paper]
  68. [arxiv 2025.9] ShortageSim: Simulating Drug Shortages under Information Asymmetry [paper]
  69. [arxiv 2025.9] Code Like Humans: A Multi-Agent Solution for Medical Coding [paper]
  70. [arxiv 2025.8] The Anatomy of a Personal Health Agent [paper]
  71. [arxiv 2025.8] MedResearcher-R1: Expert-Level Medical Deep Researcher via A Knowledge-Informed Trajectory Synthesis Framework [paper] [Github]
  72. [arxiv 2025.8] ChatThero: An LLM-Supported Chatbot for Behavior Change and Therapeutic Support in Addiction Recovery [paper]
  73. [arxiv 2025.8] Automated Clinical Problem Detection from SOAP Notes using a Collaborative Multi-Agent LLM Architecture [paper]
  74. [arxiv 2025.8] Trustworthy Agents for Electronic Health Records through Confidence Estimation [paper]
  75. [arxiv 2025.8] AT-CXR: Uncertainty-Aware Agentic Triage for Chest X-rays [paper]
  76. [arxiv 2025.8] End-to-End Agentic RAG System Training for Traceable Diagnostic Reasoning [paper]
  77. [arxiv 2025.8] Organ-Agents: Virtual Human Physiology Simulator via LLMs [paper]
  78. [arxiv 2025.8] A Multi-Agent Approach to Neurological Clinical Reasoning [paper]
  79. [arxiv 2025.8] PASS: Probabilistic Agentic Supernet Sampling for Interpretable and Adaptive Chest X-Ray Reasoning [paper]
  80. [arxiv 2025.8] HealthFlow: A Self-Evolving AI Agent with Meta Planning for Autonomous Healthcare Research [paper] [code]
  81. [arxiv 2025.8] ConfAgents: A Conformal-Guided Multi-Agent Framework for Cost-Efficient Medical Diagnosis [paper]
  82. [arxiv 2025.8] Colacare: Enhancing electronic health record modeling through large language model-driven multi-agent collaboration [paper][project page]
  83. [arxiv 2025.8] FEAT: A Multi-Agent Forensic AI System with Domain-Adapted Large Language Model for Automated Cause-of-Death Analysis [paper]
  84. [arxiv 2025.8] Are Large Language Models Dynamic Treatment Planners? An In Silico Study from a Prior Knowledge Injection Angle [paper]
  85. [arxiv 2025.8] Tree-of-Reasoning: Towards Complex Medical Diagnosis via Multi-Agent Reasoning with Evidence Tree [paper]
  86. [arxiv 2025.8] A Multi-Agent System for Complex Reasoning in Radiology Visual Question Answering [paper]
  87. [arxiv 2025.8] Patho-AgenticRAG: Towards Multimodal Agentic Retrieval-Augmented Generation for Pathology VLMs via Reinforcement Learning [paper] [code]
  88. [arxiv 2025.8] Agent-Based Feature Generation from Clinical Notes for Outcome Prediction [paper]
  89. [arxiv 2025.8] GMAT: Grounded Multi-Agent Clinical Description Generation for Text Encoder in Vision-Language MIL for Whole Slide Image Classification [paper]
  90. [arxiv 2025.8] A Multi-Agent Approach to Neurological Clinical Reasoning [paper]
  91. [biorxiv 2025.8] BioScientistAgent: Designing LLM-Biomedical Agents with KG-Augmented RL Reasoning Modules for Drug Repurposing and Mechanistic of Action Elucidation [paper]
  92. [arxiv 2025.7] Agentic AI framework for end-to-end medical data inference [paper]
  93. [arxiv 2025.7] Resilient Multi-Agent Negotiation for Medical Supply Chains: Integrating LLMs and Blockchain for Transparent Coordination [paper]
  94. [arxiv 2025.7] Intelligent Virtual Sonographer (IVS): Enhancing Physician-Robot-Patient Communication [paper]
  95. [arxiv 2025.7] A Comprehensive Survey of Electronic Health Record Modeling: From Deep Learning Approaches to Large Language Models [paper] [project page]
  96. [arxiv 2025.7] Infherno: End-to-end agent-based FHIR resource synthesis from free-form clinical notes [paper]
  97. [arxiv 2025.7] Multi-agent retrieval-augmented framework for evidence-based counterspeech against health misinformation [paper]
  98. [arxiv 2025.7] AI-VaxGuide: An Agentic RAG-Based LLM for Vaccination Decisions [paper]
  99. [arxiv 2025.7] Multi-Agent Reasoning for Cardiovascular Imaging Phenotype Analysis [paper]
  100. [arxiv 2025.7] DynamiCare: A Dynamic Multi-Agent Framework for Interactive and Open-Ended Medical Decision-Making [paper]
  101. [arxiv 2025.7] KERAP: A Knowledge-Enhanced Reasoning Approach for Accurate Zero-shot Diagnosis Prediction Using Multi-agent LLMs [paper]
  102. [arxiv 2025.7] STELLA: Self-Evolving LLM Agent for Biomedical Research [paper]
  103. [arxiv 2025.6] MedOrch: Medical Diagnosis with Tool-Augmented Reasoning Agents for Flexible Extensibility [paper]
  104. [arxiv 2025.6] MMedAgent-RL: Optimizing Multi-Agent Collaboration for Multimodal Medical Reasoning [paper] [Github]
  105. [arxiv 2025.6] From EHRs to Patient Pathways: Scalable Modeling of Longitudinal Health Trajectories with LLMs [paper]
  106. [arxiv 2025.6] Evidence-based diagnostic reasoning with multi-agent copilot for human pathology [paper]
  107. [arxiv 2025.6] An agentic system for rare disease diagnosis with traceable reasoning [paper] [demo]
  108. [arxiv 2025.6] Standard Applicability Judgment and Cross-jurisdictional Reasoning: A RAG-based Framework for Medical Device Compliance [paper]
  109. [arxiv 2025.6] From RAG to Agentic: Validating Islamic-Medicine Responses with LLM Agents [paper]
  110. [arxiv 2025.6] PRISM2: Unlocking Multi-Modal General Pathology AI with Clinical Dialogue [paper]
  111. [arxiv 2025.6] Tiered Agentic Oversight: A Hierarchical Multi-Agent System for Healthcare Safety [paper]
  112. [arxiv 2025.6] The Optimization Paradox in Clinical AI Multi-Agent Systems [paper]
  113. [arxiv 2025.6] AUTOCT: Automating Interpretable Clinical Trial Prediction with LLM Agents [paper]
  114. [arxiv 2025.6] AI Agents for Conversational Patient Triage: Preliminary Simulation-Based Evaluation with Real-World EHR Data [paper]
  115. [arxiv 2025.6] VChatter: Exploring Generative Conversational Agents for Simulating Exposure Therapy to Reduce Social Anxiety [paper]
  116. [arxiv 2025.6] AnnaAgent: Dynamic Evolution Agent System with Multi-Session Memory for Realistic Seeker Simulation [paper] [code]
  117. [arxiv 2025.6] ReflecTool: Towards Reflection-Aware Tool-Augmented Clinical Agents [paper]
  118. [arxiv 2025.6] RadFabric: Agentic AI System with Reasoning Capability for Radiology [Paper] [Project] |
  119. [arxiv 2025.5] CDR-Agent: Intelligent Selection and Execution of Clinical Decision Rules Using Large Language Model Agents [paper] [code]
  120. [arxiv 2025.5] BehaviorSFT: Behavioral Token Conditioning for Clinical Agents Across the Proactivity Spectrum [paper]
  121. [arxiv 2025.5] Silence is Not Consensus: Disrupting Agreement Bias in Multi-Agent LLMs via Catfish Agent for Clinical Decision Making [paper]
  122. [arxiv 2025.5] CPathAgent: An Agent-based Foundation Model for Interpretable High-Resolution Pathology Image Analysis Mimicking Pathologists' Diagnostic Logic [paper]
  123. [arxiv 2025.5] Are Vision Language Models Ready for Clinical Diagnosis? A 3D Medical Benchmark for Tumor-centric Visual Question Answering [paper] [code]
  124. [arxiv 2025.5] Beyond Correlation: Towards Causal Large Language Model Agents in Biomedicine [paper]
  125. [arxiv 2025.5] Generator-Mediated Bandits: Thompson Sampling for GenAI-Powered Adaptive Interventions [paper]
  126. [arxiv 2025.5] CT-Agent: A Multimodal-LLM Agent for 3D CT Radiology Question Answering [paper]
  127. [arxiv 2025.5] A Risk Taxonomy for Evaluating AI-Powered Psychotherapy Agents [paper]
  128. [arxiv 2025.5] MedAgentBoard: Benchmarking Multi-Agent Collaboration with Conventional Methods for Diverse Medical Tasks [paper] [project page]
  129. [arxiv 2025.5] A Multimodal Multi-Agent Framework for Radiology Report Generation [paper]
  130. [arxiv 2025.5] DoctorAgent-RL: A Multi-Agent Collaborative Reinforcement Learning System for Multi-Turn Clinical Dialogue [paper] [code]
  131. [biorxiv 2025.5] Biomni: A general-purpose biomedical ai agent [paper]
  132. [arxiv 2025.4] Llm agent swarm for hypothesis-driven drug discovery [paper]
  133. [arxiv 2025.4] Towards a HIPAA Compliant Agentic AI System in Healthcare [paper]
  134. [arxiv 2025.4] Customizing emotional support: How do individuals construct and interact with LLM-powered chatbots [paper]
  135. [arxiv 2025.4] Privacy-Preserving Operating Room Workflow Analysis using Digital Twins [paper]
  136. [arxiv 2025.4] An LLM-Driven Multi-Agent Debate System for Mendelian Diseases [paper]
  137. [arxiv 2025.4] Txgemma: Efficient and agentic llms for therapeutics [paper]
  138. [medrxiv 2025.4] TrialGenie: Empowering Clinical Trial Design with Agentic Intelligence and Real World Data [paper]
  139. [arxiv 2025.3] Operating room workflow analysis via reasoning segmentation over digital twins [paper]
  140. [arxiv 2025.3] TAMA: A Human--AI Collaborative Thematic Analysis Framework Using Multi-Agent LLMs for Clinical Interviews [paper]
  141. [arxiv 2025.3] Autonomous Radiotherapy Treatment Planning Using DOLA: A Privacy-Preserving, LLM-Based Optimization Agent [paper]
  142. [arxiv 2025.3] The Application of MATEC (Multi-AI Agent Team Care) Framework in Sepsis Care [paper]
  143. [arxiv 2025.3] MDTeamGPT: A Self-Evolving LLM-Based Multi-Agent Framework for Multi-Disciplinary Team Medical Consultation [paper]
  144. [arxiv 2025.3] RAG-KG-IL: A Multi-Agent Hybrid Framework for Reducing Hallucinations and Enhancing LLM Reasoning through RAG and Incremental Knowledge Graph Learning Integration [paper]
  145. [arxiv 2025.3] MAP: Evaluation and Multi-Agent Enhancement of Large Language Models for Inpatient Pathways [paper]
  146. [arxiv 2025.3] TxAgent: An AI agent for therapeutic reasoning across a universe of tools [paper]
  147. [arxiv 2025.3] MedAgentsBench: Benchmarking Thinking Models and Agent Frameworks for Complex Medical Reasoning [paper] [project page]
  148. [arxiv 2025.3] Towards conversational ai for disease management [paper]
  149. [arxiv 2025.3] GEMA-Score: Granular Explainable Multi-Agent Score for Radiology Report Evaluation [paper]
  150. [arxiv 2025.2] MIND: Towards Immersive Psychological Healing with Multi-Agent Inner Dialogue [paper]
  151. [arxiv 2025.2] Enhancing hepatopathy clinical trial efficiency: a secure, large language model-powered pre-screening pipeline [paper]
  152. [arxiv 2025.2] RAG-Enhanced Collaborative LLM Agents for Drug Discovery [paper]
  153. [arxiv 2025.2] Agentic Medical Knowledge Graphs Enhance Medical Question Answering: Bridging the Gap Between LLMs and Evolving Medical Knowledge [paper]
  154. [arxiv 2025.2] An LLM-Powered Agent for Physiological Data Analysis: A Case Study on PPG-based Heart Rate Estimation [paper]
  155. [arxiv 2025.2] Regulatory science innovation for generative AI and large language models in health and medicine: a global call for action [paper]
  156. [arxiv 2025.2] Cami: A counselor agent supporting motivational interviewing through state inference and topic exploration [paper]
  157. [arxiv 2025.2] MedRAX: Medical Reasoning Agent for Chest X-ray [paper] [code]
  158. [arxiv 2025.2] PathFinder: A Multi-Modal Multi-Agent System for Medical Diagnostic Decision-Making Applied to Histopathology [Paper] [project page]
  159. [arxiv 2025.2] M^3Builder: A Multi-Agent System for Automated Machine Learning in Medical Imaging [paper]
  160. [arxiv 2025.1] MedAgentBench: A Realistic Virtual EHR Environment to Benchmark Medical LLM Agents [paper] [project page]
  161. [arxiv 2025.1] AI Chatbots as Professional Service Agents: Developing a Professional Identity [paper]
  162. [arxiv 2025.1] Exploring the inquiry-diagnosis relationship with advanced patient simulators [paper] [project page]
  163. [arxiv 2025.1] MedXpertQA: Benchmarking Expert-Level Medical Reasoning and Understanding [paper] [project page]
  164. [arxiv 2025.1] AutoCBT: An Autonomous Multi-agent Framework for Cognitive Behavioral Therapy in Psychological Counseling [paper]
  165. [medrxiv 2025.1] Advancing the prediction and understanding of placebo responses in chronic back pain using large language models [paper]
  166. [Nature] Towards conversational diagnostic artificial intelligence [paper]
  167. [Nature Communications 2025] AgentMD: Empowering Language Agents for Risk Prediction with Large-Scale Clinical Tool Learning [paper]
  168. [Intelligent Medicine] Evaluating large language models and agents in healthcare: key challenges in clinical applications [paper]
  169. [npj Digital Medicine] Evaluating large language models as agents in the clinic [paper]
  170. [Nature Medicine 2025] An evaluation framework for clinical use of large language models in patient interaction tasks [paper]
  171. [Nature Communications 2025] An automated framework for assessing how well LLMs cite relevant medical references [paper]
  172. [Nature BME 2025] CRISPR-GPT for agentic automation of gene-editing experiments [paper]
  173. [Nature Methods 2025] GeneAgent: self-verification language agent for gene-set analysis using domain databases [paper]
  174. [npj Digital Medicine] CARE-AD: A Multi-Agent Large Language Model Framework for Alzheimer's Disease Prediction Using Longitudinal Clinical Notes [paper]
  175. [npj Digital Medicine] Vision-language model for report generation and outcome prediction in CT pulmonary angiogram [paper]
  176. [npj Artificial Intelligence] HealthcareAgent: Eliciting the Power of Large Language Models for Medical Consultation [paper]
  177. [Scientific Reports 2025] Democratizing cost-effective, agentic artificial intelligence to multilingual medical summarization through knowledge distillation [paper]
  178. [Scientific Reports 2025] A multi-agent system based on HNC for domain-specific machine translation [paper]
  179. [biorxiv 2025.6] HEAL-KGGen: A Hierarchical Multi-Agent LLM Framework with Knowledge Graph Enhancement for Genetic Biomarker-Based Medical Diagnosis [paper]
  180. [JAMIA 2025] Improving Large Language Model Applications in Biomedicine with Retrieval-Augmented Generation: A Systematic Review, Meta-Analysis, and Clinical Development Guidelines [paper]
  181. [JAMIA Open 2025] Conversational health agents: a personalized large language model-powered agent framework [paper]
  182. [JMIR] The Effectiveness of a Custom AI Chatbot for Type 2 Diabetes Mellitus Health Literacy: Development and Evaluation Study [paper]
  183. [JMIR Aging 2025] The PDC30 Chatbot—Development of a Psychoeducational Resource on Dementia Caregiving Among Family Caregivers: Mixed Methods Acceptability Study [paper]
  184. [JoVE] Evidence-based knowledge synthesis and hypothesis validation: Navigating biomedical knowledge bases via explainable ai and agentic systems [paper]
  185. [arxiv 2024.8] Drugagent: Multi-agent large language model-based reasoning for drug-target interaction prediction [paper]
  186. [Bioinformatics 2025] ESCARGOT: an AI agent leveraging large language models, dynamic graph of thoughts, and biomedical knowledge graphs for enhanced reasoning [paper]
  187. [Healthcare (Basel) 2025] MedScrubCrew: A Medical Multi-Agent Framework for Automating Appointment Scheduling Based on Patient-Provider Profile Resource Matching [paper]
  188. [Clinical Neurophysiology 2025] Agent-guided AI-powered interpretation and reporting of nerve conduction studies and EMG (INSPIRE) [paper]
  189. [Expert Systems with Applications 2025] A two-stage proactive dialogue generator for efficient clinical information collection using large language model [paper]
  190. [Physics in Medicine & Biology 2025] A feasibility study of automating radiotherapy planning with large language model agents [paper]
  191. [JCO 2025] A large language model (LLM)-based multi-agent framework for risk stratification and treatment recommendations in localized prostate cancer (locPCa). [paper]
  192. [ICDH] Voice-based AI Agents: Filling the Economic Gaps in Digital Health Delivery [paper]
  193. [IEEE EMBC 2025] Knowledge-infused LLM-powered conversational health agent: A case study for diabetes patients [paper]
  194. [ICLR 2025] MMed-RAG: Versatile Multimodal RAG System for Medical Vision Language Models [paper]
  195. [ACL 2025] Medical Graph RAG: Evidence-based Medical Large Language Model via Graph Retrieval-Augmented Generation [paper]
  196. [ACL Findings 2025] MAM: Modular Multi-Agent Framework for Multi-Modal Medical Diagnosis via Role-Specialized Collaboration [paper]
  197. [ACL Findings 2025] ASTRID--An Automated and Scalable TRIaD for the Evaluation of RAG-based Clinical Question Answering Systems [paper]
  198. [NAACL 2025] A Layered Debating Multi-Agent System for Similar Disease Diagnosis [paper]
  199. [NAACL 2025] Menti: Bridging medical calculator and llm agent with nested tool calling [paper]
  200. [COLING 2025] Unveiling performance challenges of large language models in low-resource healthcare: A demographic fairness perspective [paper]
  201. [ICMI 2025] An LLM-powered Socially Interactive Agent with Adaptive Facial Expressions for Conversing about Health [paper]
  202. [MICCAI 2025 (Oral)] WSI-Agents: A Collaborative Multi-Agent System for Multi-Modal Whole Slide Image Analysis [Paper] [GitHub]
  203. [MICCAI 2025] Multi-Agent Reasoning for Cardiovascular Imaging Phenotype Analysis [Paper] [GitHub]
  204. [MICCAI 2025] DentEval: Fine-tuning-Free Expert-Aligned Assessment in Dental Education via LLM Agents [Paper] [GitHub]
  205. [MICCAI 2025] CSAP-Assist: Instrument-Agent Dialogue Empowered Vision-Language Models for Collaborative Surgical Action Planning [Paper] [GitHub]
  206. [MICCAI 2025] MedAgentSim: Self-Evolving Multi-Agent Simulations for Realistic Clinical Interactions [Paper] [Github]
  207. [MICCAI 2025 workshop] AURA: A Multi-Modal Medical Agent for Understanding, Reasoning & Annotation [paper] [github]
  208. [ICT4AWE 2025] MentalRAG: Developing an Agentic Framework for Therapeutic Support Systems [paper]
  209. [MLHC 2025] Evaluation of Multi-Agent LLMs in Multidisciplinary Team Decision-Making for Challenging Cancer Cases [paper]
  210. [Journal of imaging informatics in medicine] AgentMRI: A Vison Language Model-Powered AI System for Self-regulating MRI Reconstruction with Multiple Degradations [paper]
  211. [COLM 2025] Can A Society of Generative Agents Simulate Human Behavior and Inform Public Health Policy? A Case Study on Vaccine Hesitancy [paper]

Year 2024

  1. [arxiv 2024.12] KG4Diagnosis: A Hierarchical Multi-Agent LLM Framework with Knowledge Graph Enhancement for Medical Diagnosis [paper]
  2. [arxiv 2024.12] PsyDraw: A Multi-Agent Multimodal System for Mental Health Screening in Left-Behind Children [paper]
  3. [IEEE Big Data] SurgBox: Agent-Driven Operating Room Sandbox with Surgery Copilot [paper] [code]
  4. [Bioinformatics] AI-HOPE: an AI-driven conversational agent for enhanced clinical and genomic data integration in precision medicine research [paper]
  5. [arxiv 2024.11] Wearable Intelligent Throat Enables Natural Speech in Stroke Patients with Dysarthria [paper]
  6. [arxiv 2024.11] PIORS: Personalized Intelligent Outpatient Reception based on Large Language Model with Multi-Agents Medical Scenario Simulation [paper] [project page]
  7. [arxiv 2024.10] IMAS: A Comprehensive Agentic Approach to Rural Healthcare Delivery [paper] [project page]
  8. [arxiv 2024.10] KGARevion: An AI Agent for Knowledge-Intensive Biomedical QA [paper]
  9. [arxiv 2024.10] Zodiac: A Cardiologist-Level LLM Framework for Multi-Agent Diagnostics [paper]
  10. [arxiv 2024.9] Simulated patient systems are intelligent when powered by large language model-based AI agents [paper]
  11. [arxiv 2024.9] On the limits of agency in agent-based models [paper]
  12. [arxiv 2024.9] Chatting Up Attachment: Using LLMs to Predict Adult Bonds [paper]
  13. [MLHC 2024] MALADE: Orchestration of LLM-powered Agents with Retrieval Augmented Generation for Pharmacovigilance [paper] [project page]
  14. [arxiv 2024.8] Agentic llm workflows for generating patient-friendly medical reports [paper] [project page]
  15. [arxiv 2024.7] Compeer: A generative conversational agent for proactive peer support [paper]
  16. [arxiv 2024.7] Cod, towards an interpretable medical agent using chain of diagnosis [paper]
  17. [arxiv 2024.7] Cactus: Towards psychological counseling conversations using cognitive behavioral theory [paper]
  18. [TMI] Integration of Multi-Source Medical Data for Medical Diagnosis Question Answering [paper]
  19. [ICLR 2025 Oral] Pathgen-1.6m: 1.6 million pathology image-text pairs generation through multi-agent collaboration [paper] [project page]
  20. [arxiv 2024.7] MentalAgora: A Gateway to Advanced Personalized Care in Mental Health through Multi-Agent Debating and Attribute Control [paper]
  21. [arxiv 2024.6] Exploring llm multi-agents for icd coding [paper]
  22. [arxiv 2024.12] Enhancing LLMs for Impression Generation in Radiology Reports through a Multi-Agent System [paper]
  23. [ICML 2024 AI for Science Workshop] TriageAgent: Towards Better Multi-Agents Collaborations for Large Language Model-Based Clinical Triage [paper]
  24. [KDD'24 Workshop] EHRFlow: A Large Language Model-Driven Iterative Multi-Agent Electronic Health Record Data Analysis Workflow [paper]
  25. [arxiv 2024.12] Agents on the Bench: Large Language Model Based Multi-Agent Framework for Trustworthy Digital Justice [paper]
  26. [arxiv 2024.6] Clinicallab: Aligning agents for multi-departmental clinical diagnostics in the real world [paper]
  27. [MLHS 2025] Path-RAG: Knowledge-Guided Key Region Retrieval for Open-ended Pathology Visual Question Answering [paper]
  28. [NeurIPS 2024] MEDIQ: Question-Asking LLMs and a Benchmark for Medical Information-Seeking [paper] [project page]
  29. [arxiv 2024.6] CliBench: A Multifaceted and Multigranular Evaluation of Clinical Diagnosis with LLMs [paper]
  30. [arxiv 2024.5] AgentClinic: a multimodal agent benchmark to evaluate AI in simulated clinical environments [paper]
  31. [arxiv 2024.5] Inquire, Interact, and Integrate: A Proactive Agent Collaborative Framework for Zero-Shot Multimodal Medical Reasoning [paper]
  32. [AAAI 2025 workshop AI4Research] Drugagent: Automating ai-aided drug discovery programming through llm multi-agent collaboration [paper]
  33. [arxiv 2024.5] Agent Hospital: A Simulacrum of Hospital with Evolvable Medical Agents [paper]
  34. [EMNLP 2024] Ehragent: Code empowers large language models for few-shot complex tabular reasoning on electronic health records [paper]
  35. [NeurIPS 2024 Oral] Mdagents: An adaptive collaboration of llms for medical decision-making [paper] [project page]
  36. [arxiv 2024.3] Llms-based few-shot disease predictions using ehr: A novel approach combining predictive agent reasoning and critical agent instruction [paper]
  37. [npj Digital Medicine] PRISM: Patient Records Interpretation for Semantic Clinical Trial Matching using Large Language Models [paper]
  38. [arxiv 2024.1] A general-purpose AI avatar in healthcare [paper]
  39. [The Lancet Digital Health] A future role for health applications of large language models depends on regulators enforcing safety standards [paper]
  40. [npj Digital Medicine] Autonomous medical evaluation for guideline adherence of large language models [paper]
  41. [Diagn Interv Radiol 2024] Large language models in radiology: fundamentals, applications, ethical considerations, risks, and future directions [paper]
  42. [PACIFIC SYMPOSIUM ON BIOCOMPUTING 2024] A conversational agent for early detection of neurotoxic effects of medications through automated intensive observation [paper]
  43. [JAMIA Open 2024] Conversational health agents: A personalized llm-powered agent framework [paper] [project page]
  44. [JMIR 2024] Mitigating cognitive biases in clinical decision-making through multi-agent conversations using large language models: simulation study [paper]
  45. [JMIR 2024] A language model--powered simulated patient with automated feedback for history taking: Prospective study [paper]
  46. [arxiv 2024.2] Development and Testing of a Novel Large Language Model-Based Clinical Decision Support Systems for Medication Safety in 12 Clinical Specialties [paper]
  47. [IEEE SoftCOM 2024] A multi-agent architecture for privacy-preserving natural language interaction with FHIR-based electronic health records [paper]
  48. [IEEE ISDFS 2024] Llm-based framework for administrative task automation in healthcare [paper]
  49. [IEEE Access 2024] Knowledge-Routed Automatic Diagnosis With Heterogeneous Patient-Oriented Graph [paper]
  50. [EMNLP Findings 2024] MMedAgent: Learning to Use Medical Tools with Multi-modal Agent [paper]
  51. [EMNLP 2024] RULE: Reliable Multimodal RAG for Factuality in Medical Vision Language Models [paper] [Github]
  52. [ACL Findings 2024] Benchmarking large language models on communicative medical coaching: a dataset and a novel system [paper]
  53. [ACL Findings 2024] Medagents: Large language models as collaborators for zero-shot medical reasoning [paper]
  54. [AAAI 2024] PathAsst: A Generative Foundation AI Assistant towards Artificial General Intelligence of Pathology [paper] [Github]
  55. [CHI 2024] Understanding the impact of long-term memory on self-disclosure with large language model-driven chatbots for public health intervention [paper]
  56. [CHI EA 2024] Conversational AI in health: Design considerations from a Wizard-of-Oz dermatology case study with users, clinicians and a medical LLM [paper]
  57. [ACM IMWUT 2024] Talk2Care: An LLM-based Voice Assistant for Communication between Healthcare Providers and Older Adults [paper]
  58. [ArabicNLP 2024] Synthetic arabic medical dialogues using advanced multi-agent llm techniques [paper]
  59. [ECCV Workshop 2024] Medco: Medical education copilots based on a multi-agent framework [paper]
  60. [Healthcare Information 2024] A Medical Consultation System for Geriatric Disease Based on Multi-agent Architecture and Knowledge Graph [paper]
  61. [Biocomputing 2025] Using large language models for efficient cancer registry coding in the real hospital setting: A feasibility study [paper]

Year 2023

  1. [NeurIPS workshop 2023] Are we going mad? benchmarking multi-agent debate between language models for medical q&a [paper]
  2. [arxiv 2023.1] Talk2Care: Facilitating asynchronous patient-provider communication with large-language-model [paper]
  3. [AMIA Annual Symposium Proceedings] Understanding the benefits and challenges of using large language model-based conversational agents for mental well-being support [paper]
  4. [Clinical NLP 2023] DERA: enhancing large language model completions with dialog-enabled resolving agents [paper] [dataset]
  5. [JMIR] The ChatGPT (generative artificial intelligence) revolution has made artificial intelligence approachable for medical professionals [paper]
  6. [JMIR] Automated monitoring of adherence to evidenced-based clinical guideline recommendations: design and implementation study [paper]
  7. [JMIR Med Educ 2023] Using ChatGPT for clinical practice and medical education: cross-sectional survey of medical students’ and physicians’ perceptions [paper]
  8. [CHI 2023] Assertiveness-based agent communication for a personalized medicine on medical imaging diagnosis [paper]

Papers by Category


1. Doctor-facing Agents

1.1 Multi-Modal Clinical Agents

(Agents designed to process and reason over multiple data types like images, text, and structured data)

Title Venue Date Paper Link Project Page
MedSAM3: Delving into Segment Anything with Medical Concepts arXiv 2025.11 Paper Star
GitHub
AURA: A Multi-modal Medical Agent for Understanding, Reasoning & Annotation MICCAI workshop 2025.07 Paper Star
GitHub
MedAgent-Pro: Towards Evidence-based Multi-modal Medical Diagnosis via Reasoning Agentic Workflow arXiv 2025.03 Paper Star
GitHub
M^3Builder: A Multi-Agent System for Automated Machine Learning in Medical Imaging arXiv 2025.02 Paper Star
GitHub
MAM: Modular Multi-Agent Framework for Multi-Modal Medical Diagnosis via Role-Specialized Collaboration ACL 2025 Paper Star
GitHub
MedAgentSim: Self-Evolving Multi-Agent Simulations for Realistic Clinical Interactions MICCAI 2025 Paper Star
GitHub
MDAgents: An Adaptive Collaboration of LLMs for Medical Decision-Making NeurIPS (Oral) 2024 Paper Star
GitHub
MMedAgent: Learning to Use Medical Tools with Multi-modal Agent EMNLP Findings 2024 Paper Star
GitHub

1.2 Radiology Agents (CT, X-ray, MRI, etc.)

Title Venue Date Paper Link Project Page
LungNoduleAgent: A Collaborative Multi-Agent System for Precision Diagnosis of Lung Nodules AAAI 2026.1 Paper Star
GitHub
Bidirectional human-AI collaboration in brain tumour assessments improves both expert human and AI agent performance arXiv 2025.12 Paper Not Available
INFORM-CT: INtegrating LLMs and VLMs FOR Incidental Findings Management in Abdominal CT arXiv 2025.12 Paper Not Available
Radiologist Copilot: Agentic AI Assistant for Holistic Radiology Reporting with Quality Control arXiv 2025.12 Paper Not Available
A Multi-Agent System for Complex Reasoning in Radiology Visual Question Answering arXiv 2025.08 Paper Not Available
AT-CXR: Uncertainty-Aware Agentic Triage for Chest X-rays arXiv 2025.08 Paper Star
GitHub
PASS: Probabilistic Agentic Supernet Sampling for Interpretable and Adaptive Chest X-Ray Reasoning arXiv 2025.08 Paper Star
GitHub
RadFabric: Agentic AI System with Reasoning Capability for Radiology arXiv 2025.06 Paper Project
A Multimodal Multi-Agent Framework for Radiology Report Generation arXiv 2025.05 Paper Not Available
CT-Agent: A Multimodal-LLM Agent for 3D CT Radiology Question Answering arXiv 2025.05 Paper Not Available
MedRAX: Medical reasoning agent for chest x-ray ICML 2025.02 Paper Star
GitHub
Vision-language model for report generation and outcome prediction in CT pulmonary angiogram npj Digital Medicine 2025 Paper Star
GitHub
AgentMRI: A Vison Language Model-Powered AI System for Self-regulating MRI Reconstruction with Multiple Degradations Journal of imaging informatics in medicine 2025 Paper Not Available
Enhancing LLMs for Impression Generation in Radiology Reports through a Multi-Agent System arXiv 2024.12 Paper Not Available

1.3 Pathology Agents

Title Venue Date Paper Link Project Page
SurvAgent: Hierarchical CoT-Enhanced Case Banking and Dichotomy-Based Multi-Agent System for Multimodal Survival Prediction arXiv 2025.11 Paper Not Available
GMAT: Grounded Multi-Agent Clinical Description Generation for Text Encoder in Vision-Language MIL arXiv 2025.08 Paper Not Available
Patho-AgenticRAG: Towards Multimodal Agentic Retrieval-Augmented Generation for Pathology VLMs arXiv 2025.08 Paper Star
GitHub
Evidence-based diagnostic reasoning with multi-agent copilot for human pathology arXiv 2025.06 Paper Not Available
CPathAgent: An Agent-based Foundation Model for Interpretable High-Resolution Pathology Image Analysis arXiv 2025.05 Paper Not Available
PathFinder: A Multi-Modal Multi-Agent System for Medical Diagnostic Decision-Making Applied to Histopathology arXiv 2025.02 Paper project
WSI-Agents: A Collaborative Multi-Agent System for Multi-Modal Whole Slide Image Analysis MICCAI (Oral) 2025 Paper Star
GitHub
Path-RAG: Knowledge-Guided Key Region Retrieval for Open-ended Pathology Visual Question Answering MLHS 2025 Paper Star
GitHub
Pathgen-1.6m: 1.6 million pathology image-text pairs generation through multi-agent collaboration ICLR (Oral) 2024 Paper Star
GitHub
PathAsst: A Generative Foundation AI Assistant towards Artificial General Intelligence of Pathology AAAI 2024 Paper Star
GitHub

1.4 Cardiovascular Imaging

Title Venue Date Paper Link Project Page
Multi-Agent Reasoning for Cardiovascular Imaging Phenotype Analysis MICCAI 2025.07 Paper Star
GitHub

1.5 Sonography / Ultrasound

Title Venue Date Paper Link Project Page
Intelligent Virtual Sonographer (IVS): Enhancing Physician-Robot-Patient Communication arXiv 2025.07 Paper Star
GitHub

1.6 Radiotherapy

Title Venue Date Paper Link Project Page
Autonomous Radiotherapy Treatment Planning Using DOLA: A Privacy-Preserving, LLM-Based Optimization Agent arXiv 2025.03 Paper Not Available
A feasibility study of automating radiotherapy planning with large language model agents Physics in Medicine & Biology 2025 Paper Not Available

1.7 Dermatology

Title Venue Date Paper Link Project Page
Conversational AI in health: Design considerations from a Wizard-of-Oz dermatology case study with users, clinicians and a medical LLM CHI 'EA 2024 Paper Not Available

1.8 Dental Agents

Title Venue Date Paper Link Project Page
DentEval: Fine-tuning-Free Expert-Aligned Assessment in Dental Education via LLM Agents MICCAI 2025 Paper Star
GitHub

1.9 Genomics & Biomarker Agents

Title Venue Date Paper Link Project Page
Geneagent: self-verification language agent for gene-set analysis using domain databases Nature Methods 2025 Paper Star
GitHub
CRISPR-GPT for agentic automation of gene-editing experiments Nature BME 2025 Paper Star
GitHub
HEAL-KGGen: A Hierarchical Multi-Agent LLM Framework for Genetic Biomarker-Based Medical Diagnosis biorxiv 2025 Paper Star
GitHub
AI-HOPE: An AI-Driven conversational agent for enhanced clinical and genomic data integration Bioinformatics 2024.12 Paper Star
GitHub

1.10 EHR & Clinical Note Agents

Title Venue Date Paper Link Project Page
Hybrid-Code: A Privacy-Preserving, Redundant Multi-Agent Framework for Reliable Local Clinical Coding arXiv 2025.12 Paper Not Available
HARMON-E: Hierarchical Agentic Reasoning for Multimodal Oncology Notes to Extract Structured Data arXiv 2025.12 Paper Not Available
ClinNoteAgents: An LLM Multi-Agent System for Predicting and Interpreting Heart Failure 30-Day Readmission from Clinical Notes arXiv 2025.12 Paper Not Available
MedDCR: Learning to Design Agentic Workflows for Medical Coding arXiv 2025.11 Paper Not Available
OEMA: Ontology-Enhanced Multi-Agent Collaboration Framework for Zero-Shot Clinical Named Entity Recognition arXiv 2025.11 Paper Not Available
Grounded by Experience: Generative Healthcare Prediction Augmented with Hierarchical Agentic Retrieval arXiv 2025.11 Paper Not Available
Traj-CoA: Patient Trajectory Modeling via Chain-of-Agents for Lung Cancer Risk Prediction NeurIPS'25 Workshop 2025.10 Paper Not Available
Automated Clinical Problem Detection from SOAP Notes using a Collaborative Multi-Agent LLM Architecture arXiv 2025.08 Paper Not Available
SNOW: Agent-Based Feature Generation from Clinical Notes for Outcome Prediction arXiv 2025.08 Paper Project
Trustworthy Agents for Electronic Health Records through Confidence Estimation arXiv 2025.8 Paper Star
GitHub
Infherno: End-to-end agent-based FHIR resource synthesis from free-form clinical notes arXiv 2025.07 Paper Star
GitHub
From EHRs to Patient Pathways: Scalable Modeling of Longitudinal Health Trajectories with LLMs arXiv 2025.6 Paper Not Available
CARE-AD: a multi-agent large language model framework for Alzheimer’s disease prediction npj Digital Medicine 2025 Paper Star
GitHub
Colacare: Enhancing electronic health record modeling through large language model-driven multi-agent collaboration arXiv 2024.10 Paper [project]
EHRFlow: A Large Language Model-Driven Iterative Multi-Agent Electronic Health Record Data Analysis Workflow KDD'24 Workshop 2024.06 Paper Star
GitHub
A multi-agent architecture for privacy-preserving natural language interaction with FHIR-based electronic health records IEEE SoftCOM 2024 Paper Not Available

1.11 Surgical Agents

Title Venue Date Paper Link Project Page
CSAP-Assist: Instrument-Agent Dialogue Empowered Vision-Language Models for Collaborative Surgical Action Planning MICCAI 2025 Paper Star
GitHub
Privacy-Preserving Operating Room Workflow Analysis using Digital Twins arXiv 2025.4 Paper Not Available

1.12 Education Agents

Title Venue Date Paper Link Project Page
An Agentic AI Framework for Training General Practitioner Student Skills arXiv 2025.12 Paper Not Available
MedTutor-R1: Socratic Personalized Medical Teaching with Multi-Agent Simulation arXiv 2025.12 Paper Not Available
Exploring Community-Powered Conversational Agent for Health Knowledge Acquisition arXiv 2025.12 Paper Not Available

1.13 Reasoning & Multi Agent Techniques

Title Venue Date Paper Link Project Page
Multi-Aspect Knowledge-Enhanced Medical Vision-Language Pretraining with Multi-Agent Data Generation ICCV 2025 Highlight 2025.12 Paper Star
Github
Incentivizing Tool-augmented Thinking with Images for Medical Image Analysis arXiv 2025.12 Paper Not Available
AdaSearch: Balancing Parametric Knowledge and Search in Large Language Models via Reinforcement Learning arXiv 2025.12 Paper Star
Github
Multi-Agent Medical Decision Consensus Matrix System: An Intelligent Collaborative Framework for Oncology MDT Consultations arXiv 2025.12 Paper Not Available
Multi-Agent Intelligence for Multidisciplinary Decision-Making in Gastrointestinal Oncology arXiv 2025.12 Paper Not Available
DART: Leveraging Multi-Agent Disagreement for Tool Recruitment in Multimodal Reasoning arXiv 2025.12 Paper Star
Github
MCP-AI: Protocol-Driven Intelligence Framework for Autonomous Reasoning in Healthcare arXiv 2025.12 Paper Not Available
Many-to-One Adversarial Consensus: Exposing Multi-Agent Collusion Risks in AI-Based Healthcare arXiv 2025.12 Paper Not Available
Thucy: An LLM-based Multi-Agent System for Claim Verification across Relational Databases AAAI Workshop 2025.12 Paper Not Available
UCAgents: Unidirectional Convergence for Visual Evidence Anchored Multi-Agent Medical Decision-Making arXiv 2025.12 Paper Star
GitHub
KOM: A Multi-Agent Artificial Intelligence System for Precision Management of Knee Osteoarthritis (KOA) arXiv 2025.11 Paper Not Available
KRAL: Knowledge and Reasoning Augmented Learning for LLM-assisted Clinical Antimicrobial Therapy arXiv 2025.11 Paper Not Available
MedResearcher-R1: Expert-Level Medical Deep Researcher via A Knowledge-Informed Trajectory Synthesis Framework arXiv 2025.8 Paper Star
GitHub
ConfAgents: A Conformal-Guided Multi-Agent Framework for Cost-Efficient Medical Diagnosis arXiv 2025.8 Paper Star
GitHub
Tree-of-Reasoning: Towards Complex Medical Diagnosis via Multi-Agent Reasoning with Evidence Tree arXiv 2025.8 Paper Star
GitHub
End-to-End Agentic RAG System Training for Traceable Diagnostic Reasoning arXiv 2025.8 Paper Star
GitHub
A Multi-Agent Approach to Neurological Clinical Reasoning arXiv 2025.8 Paper Not Available
KERAP: A knowledge-enhanced reasoning approach for accurate zero-shot diagnosis prediction arXiv 2025.7 Paper Star
GitHub
MMedAgent-RL: Optimizing Multi-Agent Collaboration for Multimodal Medical Reasoning arXiv 2025.06 Paper Not Available
An agentic system for rare disease diagnosis with traceable reasoning arXiv 2025.6 Paper [demo]
MedOrch: Medical Diagnosis with Tool-Augmented Reasoning Agents for Flexible Extensibility arXiv 2025.6 Paper Not Available
The Optimization Paradox in Clinical AI Multi-Agent Systems arXiv 2025.6 Paper Star
GitHub
DoctorAgent-RL: A Multi-Agent Collaborative Reinforcement Learning System for Multi-Turn Clinical Dialogue arXiv 2025.5 Paper Star
GitHub
Silence is Not Consensus: Disrupting Agreement Bias in Multi-Agent LLMs via Catfish Agent for Clinical Decision Making arXiv 2025.5 Paper Not Available
MDTeamGPT: A Self-Evolving LLM-Based Multi-Agent Framework for Multi-Disciplinary Team Medical Consultation arXiv 2025.3 Paper Star
GitHub
The Application of MATEC (Multi-AI Agent Team Care) Framework in Sepsis Care arXiv 2025.3 Paper Not Available
Agentic Medical Knowledge Graphs Enhance Medical Question Answering: Bridging the Gap Between LLMs and Evolving Medical Knowledge arXiv 2025.2 Paper Star
GitHub
A Layered Debating Multi-Agent System for Similar Disease Diagnosis NAACL 2025 Paper Not Available
KG4Diagnosis: A Hierarchical Multi-Agent LLM Framework with Knowledge Graph Enhancement arXiv 2024.12 Paper Not Available
Zodiac: A Cardiologist-Level LLM Framework for Multi-Agent Diagnostics arXiv 2024.10 Paper Not Available
MedAgents: Large Language Models as Collaborators for Zero-shot Medical Reasoning ACL 2024 Findings 2023.11 Paper Star
GitHub

2. Patient-Facing Applications

2.1 Mental Health & CBT Agents

Title Venue Date Paper Link Project Page
ChatThero: An LLM-Supported Chatbot for Behavior Change and Therapeutic Support in Addiction Recovery arXiv 2025.08 Paper Star
GitHub Reproduce
VChatter: Exploring Generative Conversational Agents for Simulating Exposure Therapy to Reduce Social Anxiety arXiv 2025.06 Paper Not Available
AnnaAgent: Dynamic Evolution Agent System with Multi-Session Memory for Realistic Seeker Simulation arXiv 2025.06 Paper Star
GitHub
MIND: Towards Immersive Psychological Healing with Multi-Agent Inner Dialogue ICML workshop 2025.02 Paper Star
GitHub Reproduce
Cami: A counselor agent supporting motivational interviewing through state inference and topic exploration arXiv 2025.02 Paper Star
GitHub
Autocbt: An autonomous multi-agent framework for cognitive behavioral therapy in psychological counseling arXiv 2025.01 Paper Not Available
PsyDraw: A Multi-Agent Multimodal System for Mental Health Screening in Left-Behind Children arXiv 2024.12 Paper Star
GitHub
Cactus: Towards psychological counseling conversations using cognitive behavioral theory EMNLP Findings 2024.07 Paper Star
GitHub
MentalAgora: A Gateway to Advanced Personalized Care in Mental Health through Multi-Agent Debating arXiv 2024.07 Paper Star
GitHub
Compeer: A generative conversational agent for proactive peer support arXiv 2024.07 Paper Star
GitHub
Understanding the benefits and challenges of using large language model-based conversational agents for mental well-being support AMIA Annual Symposium Proceedings 2023.07 Paper Not Available

2.2 Clinical Communication & Intake Agents

Title Venue Date Paper Link Project Page
AI Agents for Conversational Patient Triage: Preliminary Simulation-Based Evaluation with Real-World EHR Data arXiv 2025.6 paper Not Available
A two-stage proactive dialogue generator for efficient clinical information collection Expert Systems with Applications 2025 Paper Not Available
PIORS: Personalized Intelligent Outpatient Reception based on Large Language Model with Multi-Agents Medical Scenario Simulation arXiv 2024.11 Paper Star
GitHub
A language model--powered simulated patient with automated feedback for history taking: Prospective study JMIR 2024 Paper Not Available
Conversational health agents: a personalized large language model-powered agent framework JAMIA Open 2024 Paper Star
GitHub
Talk2Care: Facilitating asynchronous patient-provider communication with large-language-model arXiv 2023.9 Paper Not Available

2.3 Screening & Personalized Care Agents

Title Venue Date Paper Link Project Page
FinAgent: An Agentic AI Framework Integrating Personal Finance and Nutrition Planning arXiv 2025.12 Paper Not Available
On-device Large Multi-modal Agent for Human Activity Recognition arXiv 2025.12 Paper Not Available
Causal Reinforcement Learning based Agent-Patient Interaction with Clinical Domain Knowledge arXiv 2025.12 Paper Not Available
AI-VaxGuide: An Agentic RAG-Based LLM for Vaccination Decisions arXiv 2025.07 Paper huggingface
A Conversational Agent for Early Detection of Neurotoxic Effects of Medications through Automated Intensive Observation PACIFIC SYMPOSIUM ON BIOCOMPUTING 2024 Paper Not Available

2.4 General-purpose Healthcare Avatars

Title Venue Date Paper Link Project Page
The Anatomy of a Personal Health Agent arXiv 2025.08 Paper Not Available
A general-purpose AI avatar in healthcare arXiv 2024.01 Paper Not Available

3. Drug Discovery & Development

Title Venue Date Paper Link Project Page
MedAI: Evaluating TxAgent's Therapeutic Agentic Reasoning in the NeurIPS CURE-Bench Competition arXiv 2025.12 Paper Benchmark & Competition
BioScientistAgent: Designing LLM-Biomedical Agents with KG-Augmented RL Reasoning Modules biorxiv 2025.08 Paper Not Available
RAG-Enhanced Collaborative LLM Agents for Drug Discovery arXiv 2025.02 Paper Not Available
Large Language Model Agent for Modular Task Execution in Drug Discovery arXiv 2025.07 Paper Star
GitHub
AUTOCT: Automating Interpretable Clinical Trial Prediction with LLM Agents arXiv 2025.06 Paper Star
GitHub
Llm agent swarm for hypothesis-driven drug discovery arXiv 2025.04 Paper Not Available
Txgemma: Efficient and agentic llms for therapeutics arXiv 2025.04 Paper Not Available
TrialGenie: Empowering Clinical Trial Design with Agentic Intelligence and Real World Data medRxiv 2025.04 Paper Not Available
TxAgent: An AI agent for therapeutic reasoning across a universe of tools arXiv 2025.03 Paper Star
GitHub
Drugagent: Automating ai-aided drug discovery programming through llm multi-agent collaboration AAAI 2025 workshop AI4Research 2024.11 Paper Star
GitHub
Drugagent: Multi-agent large language model-based reasoning for drug-target interaction prediction arXiv 2024.08 Paper Star
GitHub
PRISM: Patient Records Interpretation for Semantic Clinical Trial Matching using Large Language Models npj Digital Medicine 2024.01 Paper Not Available
MALADE: Orchestration of LLM-powered Agents with Retrieval Augmented Generation for Pharmacovigilance MLHC 2024 Paper Star
GitHub

4. Healthcare Administration & Workflow

Title Venue Date Paper Link Project Page
Fair-GNE: Generalized Nash Equilibrium-Seeking Fairness in Multiagent Healthcare Automation arXiv 2025.11 Paper Not Available
MedBuild AI: An Agent-Based Hybrid Intelligence Framework for Reshaping Agency in Healthcare Infrastructure Planning through Generative Design for Medical Architecture arXiv 2025.11 Paper Not Available
ShortageSim: Simulating Drug Shortages under Information Asymmetry arXiv 2025.09 Paper Star
GitHub
Code Like Humans: A Multi-Agent Solution for Medical Coding arXiv 2025.09 Paper Not Available
Resilient Multi-Agent Negotiation for Medical Supply Chains: Integrating LLMs and Blockchain arXiv 2025.07 Paper Not Available
Standard Applicability Judgment and Cross-jurisdictional Reasoning: A RAG-based Framework for Medical Device Compliance arXiv 2025.06 Paper Not Available
Operating room workflow analysis via reasoning segmentation over digital twins arXiv 2025.03 Paper Not Available
MedScrubCrew: A Medical Multi-Agent Framework for Automating Appointment Scheduling Healthcare (Basel) 2025 Paper Not Available
IMAS: A Comprehensive Agentic Approach to Rural Healthcare Delivery arXiv 2024.10 Paper Star
GitHub
Exploring llm multi-agents for icd coding arXiv 2024.06 Paper Not Available
Llm-based framework for administrative task automation in healthcare IEEE ISDFS 2024 Paper Not Available

5. Datasets & Benchmarks

Title Venue Date Paper Link Project Page
ClinDEF: A Dynamic Evaluation Framework for Large Language Models in Clinical Reasoning arXiv 2025.12 Paper Not Available
ReX-MLE: The Autonomous Agent Benchmark for Medical Imaging Challenges arXiv 2025.12 Paper Github
MedInsightBench: Evaluating Medical Analytics Agents Through Multi-Step Insight Discovery arXiv 2025.12 Paper Not Available
CP-Env: Evaluating Large Language Models on Clinical Pathways in a Controllable Hospital Environment arXiv 2025.12 Paper Star
GitHub
AutoMedic: An Automated Evaluation Framework for Clinical Conversational Agents with Medical Dataset Grounding arXiv 2025.12 Paper Not Available
Scalably Enhancing the Clinical Validity of a Task Benchmark with Physician Oversight arXiv 2025.12 Paper Not Available
First, do NOHARM: towards clinically safe large language models arXiv 2025.12 Paper Not Available
Medical Malice: A Dataset for Context-Aware Safety in Healthcare LLMs arXiv 2025.11 Paper Not Available
MedBench v4: A Robust and Scalable Benchmark for Evaluating Chinese Medical Language Models, Multimodal Models, and Intelligent Agents arXiv 2025.11 Paper Not Available
Are Vision Language Models Ready for Clinical Diagnosis? A 3D Medical Benchmark for Tumor-centric Visual Question Answering arXiv 2025.05 Paper Star
GitHub
MedAgentBoard: Benchmarking Multi-Agent Collaboration with Conventional Methods for Diverse Medical Tasks arXiv 2025.05 Paper Star
GitHub
MedAgentsBench: Benchmarking Thinking Models and Agent Frameworks for Complex Medical Reasoning arXiv 2025.03 Paper Star
GitHub
MedAgentBench: A Realistic Virtual EHR Environment to Benchmark Medical LLM Agents arXiv 2025.01 Paper Star
GitHub
CliBench: A Multifaceted and Multigranular Evaluation of Clinical Diagnosis with LLMs arXiv 2024.06 Paper Star
GitHub
MediQ: Question-Asking LLMs for Adaptive and Reliable Clinical Reasoning arXiv 2024.06 Paper Star
GitHub
AgentClinic: a multimodal agent benchmark to evaluate AI in simulated clinical environments arXiv 2024.05 Paper Star
GitHub
Agent Hospital: A Simulacrum of Hospital with Evolvable Medical Agents arXiv 2024.05 Paper Star
GitHub

Acknowledgement

This awesome list is maintained by a collaborative team from the University of Notre Dame, Johns Hopkins University, and Emory University. The authors of the survey paper are Gelei Xu*, Xueyang Li*, Yixiong Chen*, Yuying Duan*, Shuqing Wu*, Alexander Yu*, Ching-Hao Chiu*, Juntong Ni*, Ningzhi Tang, Toby Jia-Jun Li, Alan Yuille, Wei Jin, and Yiyu Shi (* equal contribution).

Star History

Star History Chart

About

Latest Advances on Agentic AI & AI Agents for Healthcare

Resources

Stars

Watchers

Forks

Releases

No releases published

Packages

No packages published

Contributors 6