← Discover MCPs and Agents
A
AgentAI & MLGitHub

Awesome-AI-Agents-for-Healthcare

Latest Advances on Agentic AI & AI Agents for Healthcare

Links

README

From the repo.

Awesome AI Agents for Healthcare

Awesome PRs Welcome Read Our Survey Star on GitHub

This repository is a curated list of research papers, projects, and resources related to the application of Agentic AI / AI agents for healthcare, including medical image analysis, EHR manipulation, counseling, drug discovery, patient dialogue, and healthcare administration. AI agents refer to artificial intelligence systems that can autonomously perform tasks, make decisions, and interact with their environment, often through the use of large language models (LLMs), multi-agent systems, and tool integrations.

  1. The image below introduces a comprehensive conceptual framework. It provides a holistic view, detailing the pipeline from initial data perception and foundational agent capabilities to a hierarchical application ecosystem.

Overall Landscape

  1. We conducted a quantitative analysis of recent academic literature, with the key findings summarised in the image below. This analysis provides a data-driven snapshot of the field’s growth trajectory, technological underpinnings, and application hotspots:
  • Top Data Modalities: Textual data remains the most frequently utilized modality. Time-Series and Genomics exhibit a high proportion of publications from 2025.
  • Top Technologies: The technological focus is heavily concentrated on three topics: 1) developing highlevel Frameworks, 2) enhancing agent Reasoning, and 3) designing Multi-Agent collaboration paradigms.
  • Top Application Domains: Agentic AI continues to be widely applied in broad domains like General Medicine, Public Health, and Mental Health. Drug Discovery and Genomics are particularly new frontiers.

Statistics for Research Trends

We will try to keep this list updated. If you find any errors or any missing papers, please don't hesitate to open issues or pull requests.

📘 Read our survey paper here: A Comprehensive Survey of AI Agents in Healthcare

If you find our paper and repository helpful, please cite:

@article{xu2026comprehensive,
  title={A comprehensive survey of AI Agents in Healthcare},
  author={Xu, Gelei and Li, Xueyang and Chen, Yixiong and Duan, Yuying and Wu, Shuqing and Yu, Haoxinran and Chiu, Ching-Hao and Ni, Juntong and Tang, Ningzhi and Li, Toby Jia-Jun and others},
  journal={Journal of Biomedical Informatics},
  pages={105045},
  year={2026},
  publisher={Elsevier}
}

Table of Contents


Latest Papers

Year 2026

  1. [arxiv 2026.8] MARC v1: An Open-Source Multi-Agent Framework for Clinical AI Reasoning and Coordination [paper] [Github]
  2. [arxiv 2026.8] Auditable agentic AI for evidence-grounded thyroid ultrasound diagnosis and reporting [paper]
  3. [arxiv 2026.8] Social Chain of Thought: A Multi-Agent Architecture Grounded in Medical Differential Diagnosis Methodology [paper]
  4. [arxiv 2026.8] Beyond Relevance: Bayesian Evidence Acquisition for Agentic Whole-Slide Image Reasoning [paper] [Github]
  5. [arxiv 2026.8] MIRA: Medical Image Reflection for Agentic Diagnosis [paper]
  6. [arxiv 2026.8] Towards Expert-level Medical AI for Real-time Video Consultations [paper]
  7. [arxiv 2026.8] An Agentic Generative Large Language Model for Treatment Planning of Colorectal Cancer [paper]
  8. [arxiv 2026.8] Diagnosing as Cardiologists Do: ECG Agents with Doctor-Grounded Priors for Clinical Reasoning Across Diseases and Populations [paper]
  9. [arxiv 2026.8] ELICITED: EHR-grounded Longitudinal Interactive Conversations for Information-seeking Triage Evaluation and Decision-making [paper]
  10. [arxiv 2026.8] From Single Chatbots to Governed Agent Ecosystems: An Agentic AI Pattern Catalogue and Orchestration Framework for Mission-Critical Hospital Information Management Systems [paper]
  11. [arxiv 2026.8] Does Splitting a Triage Decision Across Agents Hide Bias or Help Catch It? A Multi-Agent Simulation Study of LLM-Based Resource Allocation Under Audit Capacity Constraints [paper] [Github]
  12. [arxiv 2026.8] From Siloed Algorithms to Compliance-First Agentic Platforms: A Multi-Layered Architecture for Hospital AI Systems [paper]
  13. [arxiv 2026.8] DoctorAgents: an agentic framework to iteratively refine AutoML pipeline for small clinical temporal data [paper]
  14. [arxiv 2026.8] CASCADE: An Agentic Regulatory Network Framework for Patient-Data-Validated Downstream Perturbation Prediction [paper]
  15. [arxiv 2026.8] Guideline-as-Oracle: Zero-Annotation Training of an Ophthalmic Telephone Triage Agent [paper]
  16. [arxiv 2026.8] ODRA: Synthesizing Cognitive Behavioral Therapy Sessions with Structured Chain-Of-Thought and Dynamic Patient Resistance [paper]
  17. [arxiv 2026.8] Agents Catching Agents: Shortcut Cascades and Benchmark Gaming in Clinical Multi-Agent Systems [paper] [Github]
  18. [arxiv 2026.8] TumorBoard: Evidence-Grounded Multi-Agent Decision Support for Longitudinal Neuro-Oncology [paper]
  19. [arxiv 2026.7] CyberNeuro: A Privacy-Preserving Agentic Workbench for Cohort-Scale Neuroimage and Clinical Data Analysis [paper]
  20. [CVPR 2026 Workshop] Inference-Time Agentic Decision Rules Beat Longer Evolving Search for Multi-Image Medical Reasoning [paper]
  21. [arxiv 2026.7] ClinLens: Towards Long-Horizon Coding Agents for Longitudinal Multimodal Clinical Data Science [paper]
  22. [arxiv 2026.7] Agentic AI in medicine: architectures, applications, evaluation, and challenges for clinical translation [paper]
  23. [arxiv 2026.7] PatientAgentBench: A Benchmark Framework for Evaluating Patient-Facing Health AI Agents [paper] [Github]
  24. [arxiv 2026.7] Cardiologent: Multi-Agent Clinical Decision Support for Patient-Level Arrhythmia Assessment, Urgency, and Management [paper]
  25. [arxiv 2026.7] Spectral Dynamics of Semantic Drift in Clinical Multi-Agent Language Model Networks [paper]
  26. [MICCAI 2026] Agent-Guided Relational Concept Discovery: Toward Interpretable Surgical Margin Assessment [paper]
  27. [arxiv 2026.7] Bayesian uncertainty estimation improves clinical decision making in medical AI agents [paper]
  28. [arxiv 2026.7] Harnessing Disagreement: Detecting Correlated Agreement Blindness in Multi-Agent Triage [paper]
  29. [arxiv 2026.7] MedDDC-Eval: Diagnosis-Decoupled Evaluation of Multi-Turn Medical Consultation Agents [paper]
  30. [arxiv 2026.7] Understanding From Human Perspective: A Multi-agent System for Interactive Egocentric Medical Image Segmentation [paper] [Github]
  31. [arxiv 2026.7] Cura 1T: Specialized Model for Agentic Healthcare [paper] [Github]
  32. [arxiv 2026.7] Evaluating Frontier AI Agents as Autonomous Clinical Security Auditors [paper]
  33. [arxiv 2026.7] A Multi-Agent System for Autonomous, Fine-Tuning-Free Clinical Symptom Detection: Development and Validation Study [paper]
  34. [arxiv 2026.7] Agentic systems for breast cancer treatment recommendations [paper] [Github]
  35. [arxiv 2026.7] The Path to Self-Evolving Clinical Systems: Scaling Medical Agents from Assistance to Autonomy [paper] [Github]
  36. [arxiv 2026.7] Policy-Driven CT-Agent: Modeling Phase-Aware Diagnostic Control for Clinically Consistent CT Reasoning [paper]
  37. [arxiv 2026.7] Towards Autonomous and Auditable Medical Imaging Model Development [paper]
  38. [arxiv 2026.7] Information-seeking failures of large language models in agentic clinical reasoning [paper]
  39. [arxiv 2026.7] LongMedBench: Benchmarking Medical Agents for Long-Horizon Clinical Decision-Making [paper]
  40. [arxiv 2026.7] Trust but Verify:Evidence-Linked Multi-Agent Clinical Information Extraction in Pathology [paper]
  41. [arxiv 2026.7] Toward Trustworthy Large Language Model Agents in Healthcare [paper] [Github]
  42. [arxiv 2026.7] Multi-Large Language Model Orchestrated Severity Assessment of Clinical Records (MOSAIC) [paper]
  43. [arxiv 2026.7] CogRad: A Cognitively-Inspired Multi-Agent Framework for Radiology Report Generation [paper]
  44. [arxiv 2026.7] MedCalc-Pro: Solving Complex Medical Calculations with LLM Agents [paper]
  45. [arxiv 2026.7] Evaluating Agentic Harness Systems for Autonomous Computational Pathology [paper]
  46. [arxiv 2026.6] HealthAgentBench: A Unified Benchmark Suite of Realistic Agentic Healthcare Environments for Challenging Frontier AI Agents [paper] [Github]
  47. [AMIA 2026] Agentic AI Enhances Physician Trust in Clinical Decision Making [paper]
  48. [arxiv 2026.6] TopoAgent: An Agentic Framework for Automated Topology Learning in Medical Imaging [paper]
  49. [IJCAI 2026] DEEPMED Search: An Open-Source Agentic Platform for Medical Deep Research with Introspective Verification [paper]
  50. [arxiv 2026.6] MedEvoEval: Evaluating Continual Evolution of Doctor Agents through Simulated Clinical Episodes [paper]
  51. [arxiv 2026.6] An AI agent for treatment reasoning over a biomedical tool universe [paper] [Github]
  52. [arxiv 2026.6] Why Trust Your Agent? Empirical Security Gains from TRiSM-Guided Agentic Workflows in Healthcare [paper]
  53. [MICCAI 2026] CPAgents: Agentic Composite Phenotype Generation for Cardiac Disease Association [paper]
  54. [arxiv 2026.6] Knowledge-augmented Agentic AI for Mental Health Medication Information Seeking [paper]
  55. [arxiv 2026.6] MedGuards: Multi-Agent System for Reliable Medical Error Detection and Correction [paper] [Github]
  56. [arxiv 2026.6] DeepBD: A Grounded Agentic Workflow for Variant Prioritization and Diagnosis of Genetic Birth Defects [paper]
  57. [arxiv 2026.6] EHR-Complex: Benchmarking Medical Agents for Complex Clinical Reasoning [paper]
  58. [MICCAI 2026] Evo-RAD: Navigating Rare Retinal Disease Diagnosis via Self-Evolving Agentic Retrieval [paper] [Github]
  59. [arxiv 2026.6] OpenBioRQ: Unsolved Biomedical Research Questions for Agents [paper]
  60. [arxiv 2026.6] A Multi-Agent Audit Framework for High-Stakes Reasoning: Evaluation and Interpretability in Clinical Mental Health Screening [paper]
  61. [arxiv 2026.6] BioInsight: Multi-Agent Orchestration for Interactive Biomedical Knowledge Discovery [paper]
  62. [arxiv 2026.6] Democratizing and accelerating AI-driven pathology research through agentic intelligence [paper]
  63. [arxiv 2026.6] MedRLM: Recursive Multimodal Health Intelligence for Long-Context Clinical Reasoning, Sensor-Guided Screening, Evidence-Grounded Decision Support, and Community-to-Tertiary Referral Optimization [paper]
  64. [arxiv 2026.6] Prompt, Plan, Extract: Zero-Shot Agentic LLMs Workflows for Lung Pathology Extraction from Clinical Narratives [paper]
  65. [arxiv 2026.6] Configurable Clinical Information Extraction with Agentic RAG: What Works, What Breaks, and Why [paper]
  66. [arxiv 2026.6] Are LLMs Ready to Assist Physicians? PhysAssistBench for Interactive Doctor-Patient-EHR Assistance [paper]
  67. [arxiv 2026.6] RubricsTree: Scalable and Evolving Open-Ended Evaluation of Personal Health Agents across Health Memory and Medical Skills [paper]
  68. [arxiv 2026.6] Agentic AI-based Framework for Mitigating Premature Diagnostic Handoff and Silent Hallucination in Healthcare Applications [paper]
  69. [arxiv 2026.6] MedEasy: Designing AI Standardized Patients for Clinical Consultation Training [paper]
  70. [arxiv 2026.6] Teaching agentic AI to learn expert reasoning for rare disease diagnosis [paper]
  71. [arxiv 2026.6] DeepRoot: A KG-Coordinated Multi-Agent System for Therapeutic Reasoning over Historical Medical Texts [paper]
  72. [arxiv 2026.6] Let LLMs Judge Each Other: Multi-Agent Peer-Reviewed Reasoning for Medical Question Answering [paper]
  73. [arxiv 2026.6] XMedFusion: A Knowledge-Guided Multimodal Perception and Reasoning Framework for Autonomous Medical Systems [paper]
  74. [arxiv 2026.6] Trust but Verify: Mitigating Medical Hallucinations via Post-Hoc Adversarial Auditing and Multi-Agent Feedback Loops [paper]
  75. [arxiv 2026.6] MedLatentDx: Latent Multi-Agent Communication for Cross-Hospital Rare-Disease Diagnosis [paper]
  76. [IJCAI 2026] ArogyaSutra: A Multi-Agent Framework for Multimodal Medical Reasoning in Indic Languages [paper]
  77. [arxiv 2026.6] Skill-Augmented AI Agents for Medical Research Analysis: An Exploratory Multi-Model Human Evaluation in an NSCLC Transcriptomic Biomarker Task [paper]
  78. [arxiv 2026.6] MedCTA: A Benchmark for Clinical Tool Agents [paper] [Github]
  79. [arxiv 2026.6] Experience Makes Skillful: Enabling Generalizable Medical Agent Reasoning via Self-Evolving Skill Memory [paper]
  80. [arxiv 2026.6] Baichuan-M4: A Clinical-Grade Medical Agent System for Continuous Care [paper]
  81. [arxiv 2026.6] A multi-agent system for spine MRI report generation from multi-sequence imaging [paper]
  82. [arxiv 2026.6] A Multi-modal Agentic Co-pilot for Evidence Grounded Computational Pathology [paper]
  83. [arxiv 2026.6] PSEBench: A Controllable and Verifiable Benchmark for Evaluating LLMs in Patient Safety Event Triage [paper]
  84. [arxiv 2026.6] Beyond Prompt-Based Planning: MCP-Native Graph Planning-based Biomedical Agent System [paper]
  85. [arxiv 2026.6] D2MDT: Department-aware Multidisciplinary Team Consultation with Deliberation for Efficient Clinical Prediction [paper]
  86. [arxiv 2026.6] MeDxAgent: Multi-Agent Consultation for Interactive Medical Diagnosis [paper]
  87. [arxiv 2026.6] MedCUA-Bench: A Screenshot-Only Benchmark for Clinical Computer-Use Agents [paper]
  88. [arxiv 2026.6] ClinicalMC: A Benchmark for Multi-Course Clinical Decision-Making with Large Language Models [paper]
  89. [arxiv 2026.6] Traj-Evolve: A Self-Evolving Multi-Agent System for Patient Trajectory Modeling in Lung Cancer Early Detection [paper]
  90. [arxiv 2026.6] ClinEnv: An Interactive Multi-Stage Long Horizon EHR Environment for Agents [paper]
  91. [arxiv 2026.6] AutoMedBench: Towards Medical AutoResearch with Agentic AI Models [paper]
  92. [arxiv 2026.5] SEMA-RAG: A Self-Evolving Multi-Agent Retrieval-Augmented Generation Framework for Medical Reasoning [paper]
  93. [arxiv 2026.5] CHI-Bench: Can AI Agents Automate End-to-End, Long-Horizon, Policy-Rich Healthcare Workflows? [paper] [Github] [Project]
  94. [arxiv 2026.5] COTCAgent: Preventive Consultation via Probabilistic Chain-of-Thought Completion [paper]
  95. [MICCAI 2026] DermAgent: A Self-Reflective Agentic System for Dermatological Image Analysis with Multi-Tool Reasoning and Traceable Decision-Making [paper]
  96. [arxiv 2026.5] Reinforcement Learning for Tool-Calling Agents in Fast Healthcare Interoperability Resources (FHIR) [paper]
  97. [IEEE BigData 2025] An Agentic LLM-Based Framework for Population-Scale Mental Health Screening [paper]
  98. [arxiv 2026.5] MedMemoryBench: Benchmarking Agent Memory in Personalized Healthcare [paper]
  99. [arxiv 2026.5] ABRA: Agent Benchmark for Radiology Applications [paper]
  100. [CHIL 2026] AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks [paper]
  101. [CHIL 2026] Generating synthetic electronic health record data using agent-based models to evaluate machine learning robustness under mass casualty incidents [paper]
  102. [arxiv 2026.5] DeepTumorVQA: A Hierarchical 3D CT Benchmark for Stage-Wise Evaluation of Medical VLMs and Tool-Augmented Agents [paper]
  103. [arxiv 2026.5] A Cross-Layered Multi-Drone Coordination for Medical Supply Delivery during Disaster Response Management [paper]
  104. [arxiv 2026.5] Towards Conversational Medical AI with Eyes, Ears and a Voice [paper]
  105. [arxiv 2026.5] AI-Care: A Conversational Agentic System for Task Coordination in Alzheimer's Disease Care [paper]
  106. [arxiv 2026.5] Measuring What Matters: Benchmarking Generative, Multimodal, and Agentic AI in Healthcare [paper]
  107. [AINIT 2026] Research on Security Enhancement Methods for Adversarial Robust Large Language Model Intelligent Agents for Medical Decision-Making Tasks [paper]
  108. [arxiv 2026.5] MedExAgent: Training LLM Agents to Ask, Examine, and Diagnose in Noisy Clinical Environments [paper]
  109. [arxiv 2026.5] Hygieia: A Versatile AI Agent for Rare Disease Diagnosis and Risk Gene Prioritization [paper]
  110. [arxiv 2026.5] SymptomAI: Toward a Conversational AI Agent for Everyday Symptom Assessment [paper]
  111. [arxiv 2026.5] CuraView: A Multi-Agent Framework for Medical Hallucination Detection with GraphRAG-Enhanced Knowledge Verification [paper]
  112. [arxiv 2026.5] Healthcare AI GYM for Medical Agents [paper]
  113. [arxiv 2026.5] An Empirical Study of Agent Skills for Healthcare: Practice, Gaps, and Governance [paper]
  114. [arxiv 2026.5] PhysicianBench: Evaluating LLM Agents in Real-World EHR Environments [paper]
  115. [arxiv 2026.5] GAZE: Grounded Agentic Zero-shot Evaluation with Viewer-Level Tools and Literature Retrieval on Rare Brain MRI [paper]
  116. [arxiv 2026.5] ClinicBot: A Guideline-Grounded Clinical Chatbot with Prioritized Evidence RAG and Verifiable Citations [paper]
  117. [arxiv 2026.4] Echo-α: Large Agentic Multimodal Reasoning Model for Ultrasound Interpretation [paper]
  118. [arxiv 2026.4] Detecting Clinical Discrepancies in Health Coaching Agents: A Dual-Stream Memory and Reconciliation Architecture [paper]
  119. [arxiv 2026.4] CareGuardAI: Context-Aware Multi-Agent Guardrails for Clinical Safety & Hallucination Mitigation in Patient-Facing LLMs [paper]
  120. [arxiv 2026.4] Green Shielding: A User-Centric Approach Towards Trustworthy AI [paper]
  121. [arxiv 2026.4] FastOMOP: A Foundational Architecture for Reliable Agentic Real-World Evidence Generation on OMOP CDM data [paper]
  122. [arxiv 2026.4] Vibe Medicine: Redefining Biomedical Research Through Human-AI Co-Work [paper]
  123. [arxiv 2026.4] Thinking Like a Clinician: A Cognitive AI Agent for Clinical Diagnosis via Panoramic Profiling and Adversarial Debate [paper]
  124. [ICDH IEEE 2026] Agentic AI for Personalized Physiotherapy: A Multi-Agent Framework for Generative Video Training and Real-Time Pose Correction [paper]
  125. [arxiv 2026.4] Clinically Interpretable Sepsis Early Warning via LLM-Guided Simulation of Temporal Physiological Dynamics [paper]
  126. [arxiv 2026.4] MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills [paper]
  127. [arxiv 2026.4] From Fuzzy to Formal: Scaling Hospital Quality Improvement with AI [paper]
  128. [arxiv 2026.4] Statistics, Not Scale: Modular Medical Dialogue with Bayesian Belief Engine [paper]
  129. [arxiv 2026.4] MedProbeBench: Systematic Benchmarking at Deep Evidence Integration for Expert-level Medical Guideline [paper]
  130. [arxiv 2026.4] First, Do No Harm (With LLMs): Mitigating Racial Bias via Agentic Workflows [paper]
  131. [arxiv 2026.4] Design and Evaluation of a Culturally Adapted Multimodal Virtual Agent for PTSD Screening [paper]
  132. [AAAI 2026 Bridge] Neuro-Symbolic Resolution of Recommendation Conflicts in Multimorbidity Clinical Guidelines [paper]
  133. [CSTE 2026] Persona-Based Requirements Engineering for Explainable Multi-Agent Educational Systems: A Scenario Simulator for Clinical Reasoning Training [paper]
  134. [arxiv 2026.4] Agentic Large Language Models for Training-Free Neuro-Radiological Image Analysis [paper]
  135. [ACL 2026] MARCH: Multi-Agent Radiology Clinical Hierarchy for CT Report Generation [paper]
  136. [arxiv 2026.4] DeepER-Med: Advancing Deep Evidence-Based Research in Medicine Through Agentic AI [paper]
  137. [arxiv 2026.4] RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography [paper]
  138. [arxiv 2026.4] Rethinking Patient Education as Multi-turn Multi-modal Interaction [paper]
  139. [arxiv 2026.4] Evo-MedAgent: Beyond One-Shot Diagnosis with Agents That Remember, Reflect, and Improve [paper]
  140. [arxiv 2026.4] QuarkMedSearch: A Long-Horizon Deep Search Agent for Exploring Medical Intelligence [paper]
  141. [arxiv 2026.4] VERITAS: A Multi-Agent Co-Scientist for Verifiable Image-Derived Hypothesis Testing [paper] [Github]
  142. [ACL 2026] Dialectic-Med: Mitigating Diagnostic Hallucinations via Counterfactual Adversarial Multi-Agent Debate [paper]
  143. [arxiv 2026.4] Camyla: Scaling Autonomous Research in Medical Image Segmentation [paper] [Project]
  144. [arxiv 2026.4] Sense Less, Infer More: Agentic Multimodal Transformers for Edge Medical Intelligence [paper]
  145. [IEEE ICHI 2026] BLUEmed: Retrieval-Augmented Multi-Agent Debate for Clinical Error Detection [paper]
  146. [arxiv 2026.4] HealthAdminBench: Evaluating Computer-Use Agents on Healthcare Administration Tasks [paper]
  147. [arxiv 2026.4] MedConceal: A Benchmark for Clinical Hidden-Concern Reasoning Under Partial Observability [paper]
  148. [ICLR 2026] MedVR: Annotation-Free Medical Visual Reasoning via Agentic Reinforcement Learning [paper]
  149. [ACL 2026 Findings] EMSDialog: Synthetic Multi-person Emergency Medical Service Dialogue Generation from Electronic Patient Care Reports via Multi-LLM Agents [paper]
  150. [arxiv 2026.4] Joint Optimization of Reasoning and Dual-Memory for Self-Learning Diagnostic Agent [paper]
  151. [arxiv 2026.4] LungCURE: Benchmarking Multimodal Real-World Clinical Reasoning for Precision Lung Cancer Diagnosis and Treatment [paper]
  152. [arxiv 2026.4] XrayClaw: Cooperative-Competitive Multi-Agent Alignment for Trustworthy Chest X-ray Diagnosis [paper]
  153. [arxiv 2026.4] ECG Foundation Models and Medical LLMs for Agentic Cardiovascular Intelligence at the Edge: A Review and Outlook [paper]
  154. [arxiv 2026.4] CARE: Privacy-Compliant Agentic Reasoning with Evidence Discordance [paper]
  155. [arxiv 2026.3] SkinGPT-X: A Self-Evolving Collaborative Multi-Agent System for Transparent and Trustworthy Dermatological Diagnosis [paper]
  156. [arxiv 2026.3] Symphony for Medical Coding: A Next-Generation Agentic System for Scalable and Explainable Medical Coding [paper]
  157. [arxiv 2026.3] Towards a Medical AI Scientist [paper] [Project]
  158. [arxiv 2026.3] Improving Clinical Diagnosis with Counterfactual Multi-Agent Reasoning [paper]
  159. [IEEE ICHI 2026] MediHive: A Decentralized Agent Collective for Medical Reasoning [paper]
  160. [arxiv 2026.3] Autonomous Agent-Orchestrated Digital Twins (AADT): Leveraging the OpenClaw Framework for State Synchronization in Rare Genetic Disorders [paper]
  161. [arxiv 2026.3] ClinicalAgents: Multi-Agent Orchestration for Clinical Decision Making with Dual-Memory [paper]
  162. [arxiv 2026.3] Doctorina MedBench: End-to-End Evaluation of Agent-Based Medical AI [paper]
  163. [arxiv 2026.3] Colon-Bench: An Agentic Workflow for Scalable Dense Lesion Annotation in Full-Procedure Colonoscopy Videos [paper] [Github] [Project]
  164. [arxiv 2026.3] OMIND: Framework for Knowledge Grounded Finetuning and Multi-Turn Dialogue Benchmark for Mental Health LLMs [paper]
  165. [CHI 2026 Workshop] Rethinking Health Agents: From Siloed AI to Collaborative Decision Mediators [paper]
  166. [arxiv 2026.3] MedOpenClaw: Auditable Medical Imaging Agents Reasoning over Uncurated Full Studies [paper] [Github] [Project]
  167. [arxiv 2026.3] Multi-Agent Reasoning with Consistency Verification Improves Uncertainty Calibration in Medical MCQA [paper]
  168. [CVPR 2026 Findings] CarePilot: A Multi-Agent Framework for Long-Horizon Computer Task Automation in Healthcare [paper]
  169. [ML4H 2025] Dialogue to Question Generation for Evidence-based Medical Guideline Agent Development [paper]
  170. [arxiv 2026.3] From Physician Expertise to Clinical Agents: Preserving, Standardizing, and Scaling Physicians' Medical Expertise with Lightweight LLM [paper]
  171. [arxiv 2026.3] Can LLM Agents Generate Real-World Evidence? Evaluating Observational Studies in Medical Databases [paper] [Github]
  172. [arxiv 2026.3] Cerebra: A Multidisciplinary AI Board for Multimodal Dementia Characterization and Risk Assessment [paper]
  173. [arxiv 2026.3] Agentic Automation of BT-RADS Scoring: End-to-End Multi-Agent System for Standardized Brain Tumor Follow-up Assessment [paper]
  174. [arxiv 2026.3] Unified-MAS: Universally Generating Domain-Specific Nodes for Empowering Automatic Multi-Agent Systems [paper] [Github]
  175. [ICRA 2026] Anatomical Prior-Driven Framework for Autonomous Robotic Cardiac Ultrasound Standard View Acquisition [paper]
  176. [arxiv 2026.3] Position: Multi-Agent Algorithmic Care Systems Demand Contestability for Trustworthy AI [paper]
  177. [arxiv 2026.3] Caging the Agents: A Zero Trust Security Architecture for Autonomous AI in Healthcare [paper]
  178. [arxiv 2026.3] OpenHospital: A Thing-in-itself Arena for Evolving and Benchmarking LLM-based Collective Intelligence [paper]
  179. [arxiv 2026.3] MedPriv-Bench: Benchmarking the Privacy-Utility Trade-off of Large Language Models in Medical Open-End Question Answering [paper]
  180. [arxiv 2026.3] EviAgent: Evidence-Driven Agent for Radiology Report Generation [paper]
  181. [arxiv 2026.3] Six Interventions for the Responsible and Ethical Implementation of Medical AI Agents [paper]
  182. [arxiv 2026.3] TheraAgent: Multi-Agent Framework with Self-Evolving Memory and Evidence-Calibrated Reasoning for PET Theranostics [paper]
  183. [arxiv 2026.3] When OpenClaw Meets Hospital: Toward an Agentic Operating System for Dynamic Clinical Workflows [paper]
  184. [arxiv 2026.3] UAV-MARL: Multi-Agent Reinforcement Learning for Time-Critical and Dynamic Medical Supply Delivery [paper]
  185. [arxiv 2026.3] MedMASLab: A Unified Orchestration Framework for Benchmarking Multimodal Medical Multi-Agent Systems [paper]
  186. [arxiv 2026.3] RexDrug: Reliable Multi-Drug Combination Extraction through Reasoning-Enhanced LLMs [paper] [Github]
  187. [arxiv 2026.3] YAQIN: Culturally Sensitive, Agentic AI for Mental Healthcare Support Among Muslim Women in the UK [paper]
  188. [arxiv 2026.3] Empowering Locally Deployable Medical Agent via State Enhanced Logical Skills for FHIR-based Clinical Tasks [paper]
  189. [arxiv 2026.3] Computational Pathology in the Era of Emerging Foundation and Agentic AI [paper]
  190. [arxiv 2026.3] Shifting Adaptation from Weight Space to Memory Space: A Memory-Augmented Agent for Medical Image Segmentation [paper]
  191. [arxiv 2026.3] Evolving Medical Imaging Agents via Experience-driven Self-skill Discovery [paper]
  192. [arxiv 2026.3] Meissa: Multi-modal Medical Agentic Intelligence [paper] [Github]
  193. [ICLR 2026] CARE: Towards Clinical Accountability in Multi-Modal Medical Reasoning with an Evidence-Grounded Agentic Framework [paper]
  194. [ICLR 2026] ATPO: Adaptive Tree Policy Optimization for Multi-Turn Medical Dialogue [paper]
  195. [EACL 2026 Workshop] Do Mixed-Vendor Multi-Agent LLMs Improve Clinical Diagnosis? [paper]
  196. [arxiv 2026.3] MedCoRAG: Interpretable Hepatology Diagnosis via Hybrid Evidence Retrieval and Multispecialty Consensus [paper]
  197. [arxiv 2026.3] MedCollab: Causal-Driven Multi-Agent Collaboration for Full-Cycle Clinical Diagnosis via IBIS-Structured Argumentation [paper]
  198. [arxiv 2026.3] From Conflict to Consensus: Boosting Medical Reasoning via Multi-Round Agentic RAG [paper] [Github]
  199. [arxiv 2026.3] MIND: Unified Inquiry and Diagnosis RL with Criteria Grounded Clinical Supports for Psychiatric Consultation [paper]
  200. [arxiv 2026.3] DUCX: Decomposing Unfairness in Tool-Using Chest X-ray Agents [paper]
  201. [arxiv 2026.3] OPGAgent: An Agent for Auditable Dental Panoramic X-ray Interpretation [paper]
  202. [arxiv 2026.3] TARSE: Test-Time Adaptation via Retrieval of Skills and Experience for Reasoning Agents [paper]
  203. [arxiv 2026.3] ProtRLSearch: A Multi-Round Multimodal Protein Search Agent with Large Language Models Trained via Reinforcement Learning [paper]
  204. [arxiv 2026.3] A Multi-Agent Framework for Interpreting Multivariate Physiological Time Series [paper]
  205. [HealthSec/ACSAC 2026] Goal-Driven Risk Assessment for LLM-Powered Systems: A Healthcare Case Study [paper]
  206. [arxiv 2026.2] 3DMedAgent: Unified Perception-to-Understanding for 3D Medical Analysis [paper]
  207. [MICCAI 2026] Can Agents Distinguish Visually Hard-to-Separate Diseases in a Zero-Shot Setting? [paper] [Github]
  208. [arxiv 2026.2] Which Tool Response Should I Trust? Tool-Expertise-Aware Chest X-ray Agent with Multimodal Agentic Learning [paper]
  209. [arxiv 2026.2] MedClarify: An Information-Seeking AI Agent for Medical Diagnosis with Case-Specific Follow-up Questions [paper]
  210. [arxiv 2026.2] LAMMI-Pathology: A Tool-Centric Bottom-Up LVLM-Agent Framework for Molecularly Informed Medical Intelligence in Pathology [paper]
  211. [arxiv 2026.2] NutriOrion: A Hierarchical Multi-Agent Framework for Personalized Nutrition Intervention Grounded in Clinical Guidelines [paper]
  212. [arxiv 2026.2] TRACE: Temporal Reasoning via Agentic Context Evolution for Streaming Electronic Health Records [paper]
  213. [arxiv 2026.2] CoMMa: Contribution-Aware Medical Multi-Agents From A Game-Theoretic Perspective [paper]
  214. [AAAI 2026 Workshop] SynthAgent: A Multi-Agent LLM Framework for Realistic Patient Simulation [paper]
  215. [arxiv 2026.2] MedCoG: Maximizing LLM Inference Density in Medical Reasoning via Meta-Cognitive Regulation [paper]
  216. [arxiv 2026.2] Picking the Right Specialist: Attentive Neural Process-based Selection of Task-Specialized Models as Tools for Agentic Healthcare Systems [paper]
  217. [arxiv 2026.2] A Multi-Agent Framework for Medical AI: Leveraging Fine-Tuned GPT, LLaMA, and DeepSeek R1 for Evidence-Based and Bias-Aware Clinical Query Processing [paper]
  218. [arxiv 2026.2] MedScope: Incentivizing "Think with Videos" for Clinical Reasoning via Coarse-to-Fine Tool Calling [paper]
  219. [arxiv 2026.2] MedXIAOHE: A Comprehensive Recipe for Building Medical MLLMs [paper]
  220. [arxiv 2026.2] Advancing AI Trustworthiness Through Patient Simulation: Risk Assessment of Conversational Agents for Antidepressant Selection [paper]
  221. [arxiv 2026.2] LiveMedBench: A Contamination-Free Medical Benchmark for LLMs with Automated Rubric Evaluation [paper]
  222. [arxiv 2026.2] Closing Reasoning Gaps in Clinical Agents with Differential Reasoning Learning [paper]
  223. [ICHI 2026] Human-Guided Agentic AI for Multimodal Clinical Prediction: Lessons from the AgentDS Healthcare Benchmark [paper]
  224. [arxiv 2026.2] ALPACA: A Reinforcement Learning Environment for Medication Repurposing and Treatment Optimization in Alzheimer's Disease [paper]
  225. [arxiv 2026.2] The Doctor Will (Still) See You Now: On the Structural Limits of Agentic AI in Healthcare [paper]
  226. [arxiv 2026.2] Agentic AI, Medical Morality, and the Transformation of the Patient-Physician Relationship [paper]
  227. [IEEE Access 2026] Agentic AI in Healthcare & Medicine: A Seven-Dimensional Taxonomy for Empirical Evaluation of LLM-based Agents [paper]
  228. [arxiv 2026.2] MedSAM-Agent: Empowering Interactive Medical Image Segmentation with Multi-turn Agentic Reinforcement Learning [paper] [Github]
  229. [arxiv 2026.2] Pruning Minimal Reasoning Graphs for Efficient Retrieval-Augmented Generation [paper]
  230. [arxiv 2026.2] RE-MCDF: Closed-Loop Multi-Expert LLM Reasoning for Knowledge-Grounded Clinical Diagnosis [paper]
  231. [arxiv 2026.2] MedBeads: An Agent-Native, Immutable Data Substrate for Trustworthy Medical AI [paper]
  232. [arxiv 2026.2] AutoHealth: An Uncertainty-Aware Multi-Agent System for Autonomous Health Data Modeling [paper]
  233. [CAIN 2026] Engineering AI Agents for Clinical Workflows: A Case Study in Architecture, MLOps, and Governance [paper]
  234. [arxiv 2026.2] ExperienceWeaver: Optimizing Small-sample Experience Learning for LLM-based Clinical Text Improvement [paper]
  235. [arxiv 2026.1] EvoClinician: A Self-Evolving Agent for Multi-Turn Medical Diagnosis via Test-Time Evolutionary Learning [paper] [Github]
  236. [arxiv 2026.1] Scaling Medical Reasoning Verification via Tool-Integrated Reinforcement Learning [paper]
  237. [arxiv 2026.1] DEEPMED: Building a Medical DeepResearch Agent via Multi-hop Med-Search Data [paper]
  238. [arxiv 2026.1] AgentsEval: Clinically Faithful Evaluation of Medical Imaging Reports via Multi-Agent Reasoning [paper]
  239. [arxiv 2026.1] AgentEHR: Advancing Autonomous Clinical Decision-Making via Retrospective Summarization [paper]
  240. [arxiv 2026.1] MedConsultBench: A Full-Cycle, Fine-Grained, Process-Aware Benchmark for Medical Consultation Agents [paper]
  241. [EACL 2026] Knowing When to Abstain: Medical LLMs Under Clinical Uncertainty [paper]
  242. [arxiv 2026.1] Route, Retrieve, Reflect, Repair: Self-Improving Agentic Framework for Visual Detection and Linguistic Reasoning in Medical Imaging [paper] [Github]
  243. [arxiv 2026.1] MEDVISTAGYM: A Scalable Training Environment for Thinking with Medical Images via Tool-Integrated Reinforcement Learning [paper]
  244. [arxiv 2026.1] MedEinst: Benchmarking the Einstellung Effect in Medical LLMs through Counterfactual Differential Diagnosis [paper]
  245. [arxiv 2026.1] DemMA: Dementia Multi-Turn Dialogue Agent with Expert-Guided Reasoning and Action Simulation [paper]
  246. [arxiv 2026.1] IBISAgent: Reinforcing Pixel-Level Visual Reasoning in MLLMs for Universal Biomedical Object Referring and Segmentation [paper]
  247. [arxiv 2026.1] Bayesian Orchestration of Multi-LLM Agents for Cost-Aware Sequential Decision-Making [paper]
  248. [arxiv 2026.1] An Explainable Agentic AI Framework for Uncertainty-Aware and Abstention-Enabled Acute Ischemic Stroke Imaging Decisions [paper]
  249. [AAAI 2026] ShortageSim: Simulating Drug Shortages under Information Asymmetry [paper]
  250. [ICLR 2026] MedAgentGym: Training LLM Agents for Code-Based Medical Reasoning at Scale [paper] [Github]
  251. [AAAI 2026] LungNoduleAgent: A Collaborative Multi-Agent System for Precision Diagnosis of Lung Nodules [paper] [Github]
  252. [Nature Communications 2026] Wearable Intelligent Throat Enables Natural Speech in Stroke Patients with Dysarthria [paper]
  253. [npj Artificial Intelligence 2026] AI agent in healthcare: applications, evaluations, and future directions [paper]
  254. [npj Digital Medicine 2026] Benchmarking large language model-based agent systems for clinical decision tasks [paper]
  255. [npj Digital Medicine 2026] Reimagining psychiatric care with agentic AI: promise, challenges, and a roadmap forward [paper]
  256. [Nature Biotechnology 2026] Agentic AI and the rise of in silico team science in biomedical research [paper]

Year 2025

  1. [arxiv 2025.12] Hybrid-Code: A Privacy-Preserving, Redundant Multi-Agent Framework for Reliable Local Clinical Coding [paper]
  2. [arxiv 2025.12] ClinDEF: A Dynamic Evaluation Framework for Large Language Models in Clinical Reasoning [paper]
  3. [arxiv 2025.12] HARMON-E: Hierarchical Agentic Reasoning for Multimodal Oncology Notes to Extract Structured Data [paper]
  4. [arxiv 2025.12] Bidirectional human-AI collaboration in brain tumour assessments improves both expert human and AI agent performance [paper]
  5. [arxiv 2025.12] On-device Large Multi-modal Agent for Human Activity Recognition [paper]
  6. [arxiv 2025.12] Scalably Enhancing the Clinical Validity of a Task Benchmark with Physician Oversight [paper]
  7. [arxiv 2025.12] Agent-Based Output Drift Detection for Breast Cancer Response Prediction in a Multisite Clinical Decision Support System [paper]
  8. [arxiv 2025.12] An Agentic AI Framework for Training General Practitioner Student Skills [paper]
  9. [arxiv 2025.12] ReX-MLE: The Autonomous Agent Benchmark for Medical Imaging Challenges [paper]
  10. [arxiv 2025.12] AdaSearch: Balancing Parametric Knowledge and Search in Large Language Models via Reinforcement Learning [paper]
  11. [arxiv 2025.12] A Multi-Agent Large Language Model Framework for Automated Qualitative Analysis [paper]
  12. [arxiv 2025.12] Mapis: A Knowledge-Graph Grounded Multi-Agent Framework for Evidence-Based PCOS Diagnosis [paper]
  13. [arxiv 2025.12] INFORM-CT: INtegrating LLMs and VLMs FOR Incidental Findings Management in Abdominal CT [paper]
  14. [arxiv 2025.12] Multi-Agent Medical Decision Consensus Matrix System: An Intelligent Collaborative Framework for Oncology MDT Consultations [paper]
  15. [arxiv 2025.12] Incentivizing Tool-augmented Thinking with Images for Medical Image Analysis [paper]
  16. [arxiv 2025.12] MedInsightBench: Evaluating Medical Analytics Agents Through Multi-Step Insight Discovery in Multimodal Medical Data [paper]
  17. [arxiv 2025.12] Socratic Students: Teaching Language Models to Learn by Asking Questions [paper]
  18. [arxiv 2025.12] MedAI: Evaluating TxAgent's Therapeutic Agentic Reasoning in the NeurIPS CURE-Bench Competition [paper] [Benchmark & Competition]
  19. [arxiv 2025.12] CP-Env: Evaluating Large Language Models on Clinical Pathways in a Controllable Hospital Environment [paper] [Github]
  20. [arxiv 2025.12] AutoMedic: An Automated Evaluation Framework for Clinical Conversational Agents with Medical Dataset Grounding [paper]
  21. [arxiv 2025.12] Exploring Community-Powered Conversational Agent for Health Knowledge Acquisition: A Case Study in Colorectal Cancer [paper]
  22. [arxiv 2025.12] Multi-Agent Intelligence for Multidisciplinary Decision-Making in Gastrointestinal Oncology [paper]
  23. [arxiv 2025.12] DART: Leveraging Multi-Agent Disagreement for Tool Recruitment in Multimodal Reasoning [paper]
  24. [arxiv 2025.12] ClinNoteAgents: An LLM Multi-Agent System for Predicting and Interpreting Heart Failure 30-Day Readmission from Clinical Notes [paper]
  25. [arxiv 2025.12] MedTutor-R1: Socratic Personalized Medical Teaching with Multi-Agent Simulation [paper] [Github]
  26. [arxiv 2025.12] MCP-AI: Protocol-Driven Intelligence Framework for Autonomous Reasoning in Healthcare [paper]
  27. [arxiv 2025.12] Multi-Aspect Knowledge-Enhanced Medical Vision-Language Pretraining with Multi-Agent Data Generation [paper]
  28. [arxiv 2025.12] Thucy: An LLM-based Multi-Agent System for Claim Verification across Relational Databases [paper]
  29. [arxiv 2025.12] Many-to-One Adversarial Consensus: Exposing Multi-Agent Collusion Risks in AI-Based Healthcare [paper]
  30. [arxiv 2025.12] FinAgent: An Agentic AI Framework Integrating Personal Finance and Nutrition Planning [paper]
  31. [arxiv 2025.12] Radiologist Copilot: Agentic AI Assistant for Holistic Radiology Reporting with Quality Control [paper]
  32. [arxiv 2025.12] UCAgents: Unidirectional Convergence for Visual Evidence Anchored Multi-Agent Medical Decision-Making [paper]
  33. [arxiv 2025.12] First, do NOHARM: towards clinically safe large language models [paper]
  34. [arxiv 2025.12] Causal Reinforcement Learning based Agent-Patient Interaction with Clinical Domain Knowledge [paper]
  35. [arxiv 2025.11] MedEyes: Learning Dynamic Visual Focus for Medical Progressive Diagnosis [paper] [GitHub]
  36. [arxiv 2025.11] MedSAM3: Delving into Segment Anything with Medical Concepts [paper] [Github]
  37. [arxiv 2025.11] SurvAgent: Hierarchical CoT-Enhanced Case Banking and Dichotomy-Based Multi-Agent System for Multimodal Survival Prediction [paper]
  38. [arxiv 2025.11] KOM: A Multi-Agent Artificial Intelligence System for Precision Management of Knee Osteoarthritis (KOA) [paper]
  39. [arxiv 2025.11] KRAL: Knowledge and Reasoning Augmented Learning for LLM-assisted Clinical Antimicrobial Therapy [paper]
  40. [arxiv 2025.11] Medical Malice: A Dataset for Context-Aware Safety in Healthcare LLMs [paper]
  41. [arxiv 2025.11] MedBench v4: A Robust and Scalable Benchmark for Evaluating Chinese Medical Language Models, Multimodal Models, and Intelligent Agents [paper]
  42. [arxiv 2025.11] Fair-GNE: Generalized Nash Equilibrium-Seeking Fairness in Multiagent Healthcare Automation [paper]
  43. [arxiv 2025.11] MedDCR: Learning to Design Agentic Workflows for Medical Coding [paper]
  44. [arxiv 2025.11] Grounded by Experience: Generative Healthcare Prediction Augmented with Hierarchical Agentic Retrieval [paper]
  45. [arxiv 2025.11] OEMA: Ontology-Enhanced Multi-Agent Collaboration Framework for Zero-Shot Clinical Named Entity Recognition [paper]
  46. [arxiv 2025.11] MedBuild AI: An Agent-Based Hybrid Intelligence Framework for Reshaping Agency in Healthcare Infrastructure Planning through Generative Design for Medical Architecture [paper]
  47. [arxiv 2025.11] From Passive to Proactive: A Multi-Agent System with Dynamic Task Orchestration for Intelligent Medical Pre-Consultation [paper]
  48. [arxiv 2025.11] Fine-Tuning DialoGPT on Common Diseases in Rural Nepal for Medical Conversations [paper]
  49. [arxiv 2025.10] Traj-CoA: Patient Trajectory Modeling via Chain-of-Agents for Lung Cancer Risk Prediction [paper]
  50. [arxiv 2025.10] FT-ARM: Fine-Tuned Agentic Reflection Multimodal Language Model for Pressure Ulcer Severity Classification with Reasoning [paper]
  51. [arxiv 2025.10] SNOMED CT-powered Knowledge Graphs for Structured Clinical Data and Diagnostic Reasoning [paper]
  52. [arxiv 2025.10] Speculative Model Risk in Healthcare AI: Using Storytelling to Surface Unintended Harms [paper]
  53. [arxiv 2025.10] MedCoAct: Confidence-Aware Multi-Agent Collaboration for Complete Clinical Decision [paper]
  54. [arxiv 2025.10] Haibu Mathematical-Medical Intelligent Agent:Enhancing Large Language Model Reliability in Medical Tasks via Verifiable Reasoning Chains [paper]
  55. [arxiv 2025.10] Reinforcement Learning for Clinical Reasoning: Aligning LLMs with ACR Imaging Appropriateness Criteria [paper]
  56. [EMNLP 2025 Industry] CLARITY: Clinical Assistant for Routing, Inference, and Triage [paper]
  57. [arxiv 2025.10] Secure Multi-Modal Data Fusion in Federated Digital Health Systems via MCP [paper]
  58. [arxiv 2025.9] AgenticAD: A Specialized Multiagent System Framework for Holistic Alzheimer Disease Management [paper]
  59. [arxiv 2025.9] Agentic-AI Healthcare: Multilingual, Privacy-First Framework with MCP Agents [paper]
  60. [arxiv 2025.9] Online Decision Making with Generative Action Sets [paper]
  61. [arxiv 2025.9] PAME-AI: Patient Messaging Creation and Optimization using Agentic AI [paper]
  62. [arxiv 2025.9] ToolUniverse: An open platform for democratizing AI scientists [paper] [Github]
  63. [arxiv 2025.9] A co-evolving agentic AI system for medical imaging analysis [paper]
  64. [arxiv 2025.9] FHIR-AgentBench: Benchmarking LLM Agents for Realistic Interoperable EHR Question Answering [paper] [Github]
  65. [arxiv 2025.9] MedFact: Benchmarking the Fact-Checking Capabilities of Large Language Models on Chinese Medical Texts [paper] [Github]
  66. [arxiv 2025.9] Agentic Temporal Graph of Reasoning with Multimodal Language Models: A Potential AI Aid to Healthcare [paper]
  67. [arxiv 2025.9] Using AI to Optimize Patient Transfer and Resource Utilization During Mass-Casualty Incidents: A Simulation Platform [paper]
  68. [arxiv 2025.9] Demo: Healthcare Agent Orchestrator (HAO) for Patient Summarization in Molecular Tumor Boards [paper] [Github]
  69. [arxiv 2025.9] Chatbot To Help Patients Understand Their Health [paper]
  70. [arxiv 2025.9] Code Like Humans: A Multi-Agent Solution for Medical Coding [paper]
  71. [arxiv 2025.8] The Anatomy of a Personal Health Agent [paper]
  72. [arxiv 2025.8] MedResearcher-R1: Expert-Level Medical Deep Researcher via A Knowledge-Informed Trajectory Synthesis Framework [paper] [Github]
  73. [arxiv 2025.8] ChatThero: An LLM-Supported Chatbot for Behavior Change and Therapeutic Support in Addiction Recovery [paper]
  74. [arxiv 2025.8] Automated Clinical Problem Detection from SOAP Notes using a Collaborative Multi-Agent LLM Architecture [paper]
  75. [arxiv 2025.8] Trustworthy Agents for Electronic Health Records through Confidence Estimation [paper]
  76. [arxiv 2025.8] AT-CXR: Uncertainty-Aware Agentic Triage for Chest X-rays [paper]
  77. [arxiv 2025.8] End-to-End Agentic RAG System Training for Traceable Diagnostic Reasoning [paper]
  78. [arxiv 2025.8] Organ-Agents: Virtual Human Physiology Simulator via LLMs [paper]
  79. [arxiv 2025.8] A Multi-Agent Approach to Neurological Clinical Reasoning [paper]
  80. [arxiv 2025.8] PASS: Probabilistic Agentic Supernet Sampling for Interpretable and Adaptive Chest X-Ray Reasoning [paper]
  81. [arxiv 2025.8] HealthFlow: A Self-Evolving AI Agent with Meta Planning for Autonomous Healthcare Research [paper] [code]
  82. [arxiv 2025.8] ConfAgents: A Conformal-Guided Multi-Agent Framework for Cost-Efficient Medical Diagnosis [paper]
  83. [arxiv 2025.8] Colacare: Enhancing electronic health record modeling through large language model-driven multi-agent collaboration [paper][project page]
  84. [arxiv 2025.8] FEAT: A Multi-Agent Forensic AI System with Domain-Adapted Large Language Model for Automated Cause-of-Death Analysis [paper]
  85. [arxiv 2025.8] Are Large Language Models Dynamic Treatment Planners? An In Silico Study from a Prior Knowledge Injection Angle [paper]
  86. [arxiv 2025.8] Tree-of-Reasoning: Towards Complex Medical Diagnosis via Multi-Agent Reasoning with Evidence Tree [paper]
  87. [arxiv 2025.8] A Multi-Agent System for Complex Reasoning in Radiology Visual Question Answering [paper]
  88. [arxiv 2025.8] Patho-AgenticRAG: Towards Multimodal Agentic Retrieval-Augmented Generation for Pathology VLMs via Reinforcement Learning [paper] [code]
  89. [arxiv 2025.8] Agent-Based Feature Generation from Clinical Notes for Outcome Prediction [paper]
  90. [arxiv 2025.8] GMAT: Grounded Multi-Agent Clinical Description Generation for Text Encoder in Vision-Language MIL for Whole Slide Image Classification [paper]
  91. [biorxiv 2025.8] BioScientistAgent: Designing LLM-Biomedical Agents with KG-Augmented RL Reasoning Modules for Drug Repurposing and Mechanistic of Action Elucidation [paper]
  92. [arxiv 2025.7] Agentic AI framework for end-to-end medical data inference [paper]
  93. [arxiv 2025.7] Resilient Multi-Agent Negotiation for Medical Supply Chains: Integrating LLMs and Blockchain for Transparent Coordination [paper]
  94. [arxiv 2025.7] Intelligent Virtual Sonographer (IVS): Enhancing Physician-Robot-Patient Communication [paper]
  95. [arxiv 2025.7] A Comprehensive Survey of Electronic Health Record Modeling: From Deep Learning Approaches to Large Language Models [paper] [project page]
  96. [arxiv 2025.7] Infherno: End-to-end agent-based FHIR resource synthesis from free-form clinical notes [paper]
  97. [arxiv 2025.7] Multi-agent retrieval-augmented framework for evidence-based counterspeech against health misinformation [paper]
  98. [arxiv 2025.7] AI-VaxGuide: An Agentic RAG-Based LLM for Vaccination Decisions [paper]
  99. [arxiv 2025.7] Multi-Agent Reasoning for Cardiovascular Imaging Phenotype Analysis [paper]
  100. [arxiv 2025.7] DynamiCare: A Dynamic Multi-Agent Framework for Interactive and Open-Ended Medical Decision-Making [paper]
  101. [arxiv 2025.7] KERAP: A Knowledge-Enhanced Reasoning Approach for Accurate Zero-shot Diagnosis Prediction Using Multi-agent LLMs [paper]
  102. [arxiv 2025.7] STELLA: Self-Evolving LLM Agent for Biomedical Research [paper] [Github]
  103. [arxiv 2025.6] MedOrch: Medical Diagnosis with Tool-Augmented Reasoning Agents for Flexible Extensibility [paper]
  104. [arxiv 2025.6] MMedAgent-RL: Optimizing Multi-Agent Collaboration for Multimodal Medical Reasoning [paper]
  105. [arxiv 2025.6] From EHRs to Patient Pathways: Scalable Modeling of Longitudinal Health Trajectories with LLMs [paper]
  106. [arxiv 2025.6] Evidence-based diagnostic reasoning with multi-agent copilot for human pathology [paper]
  107. [arxiv 2025.6] An agentic system for rare disease diagnosis with traceable reasoning [paper] [demo]
  108. [arxiv 2025.6] Standard Applicability Judgment and Cross-jurisdictional Reasoning: A RAG-based Framework for Medical Device Compliance [paper]
  109. [arxiv 2025.6] From RAG to Agentic: Validating Islamic-Medicine Responses with LLM Agents [paper]
  110. [arxiv 2025.6] PRISM2: Unlocking Multi-Modal General Pathology AI with Clinical Dialogue [paper]
  111. [arxiv 2025.6] Tiered Agentic Oversight: A Hierarchical Multi-Agent System for Healthcare Safety [paper]
  112. [arxiv 2025.6] The Optimization Paradox in Clinical AI Multi-Agent Systems [paper]
  113. [EMNLP 2025] AUTOCT: Automating Interpretable Clinical Trial Prediction with LLM Agents [paper]
  114. [arxiv 2025.6] AI Agents for Conversational Patient Triage: Preliminary Simulation-Based Evaluation with Real-World EHR Data [paper]
  115. [arxiv 2025.6] VChatter: Exploring Generative Conversational Agents for Simulating Exposure Therapy to Reduce Social Anxiety [paper]
  116. [ACL 2025 Findings] AnnaAgent: Dynamic Evolution Agent System with Multi-Session Memory for Realistic Seeker Simulation [paper] [code]
  117. [ACL 2025] ReflecTool: Towards Reflection-Aware Tool-Augmented Clinical Agents [paper] [Github] [Project]
  118. [arxiv 2025.6] RadFabric: Agentic AI System with Reasoning Capability for Radiology [Paper] [Project] |
  119. [arxiv 2025.5] CDR-Agent: Intelligent Selection and Execution of Clinical Decision Rules Using Large Language Model Agents [paper] [code]
  120. [arxiv 2025.5] BehaviorSFT: Behavioral Token Conditioning for Clinical Agents Across the Proactivity Spectrum [paper]
  121. [arxiv 2025.5] Silence is Not Consensus: Disrupting Agreement Bias in Multi-Agent LLMs via Catfish Agent for Clinical Decision Making [paper]
  122. [NeurIPS 2025] CPathAgent: An Agent-based Foundation Model for Interpretable High-Resolution Pathology Image Analysis Mimicking Pathologists' Diagnostic Logic [paper]
  123. [arxiv 2025.5] Are Vision Language Models Ready for Clinical Diagnosis? A 3D Medical Benchmark for Tumor-centric Visual Question Answering [paper] [code]
  124. [arxiv 2025.5] Beyond Correlation: Towards Causal Large Language Model Agents in Biomedicine [paper]
  125. [NeurIPS 2025] Generator-Mediated Bandits: Thompson Sampling for GenAI-Powered Adaptive Interventions [paper]
  126. [arxiv 2025.5] CT-Agent: A Multimodal-LLM Agent for 3D CT Radiology Question Answering [paper]
  127. [arxiv 2025.5] A Risk Taxonomy for Evaluating AI-Powered Psychotherapy Agents [paper]
  128. [NeurIPS 2025] MedAgentBoard: Benchmarking Multi-Agent Collaboration with Conventional Methods for Diverse Medical Tasks [paper] [project page]
  129. [arxiv 2025.5] A Multimodal Multi-Agent Framework for Radiology Report Generation [paper]
  130. [EMNLP 2025] DoctorAgent-RL: A Multi-Agent Collaborative Reinforcement Learning System for Multi-Turn Clinical Dialogue [paper] [code]
  131. [biorxiv 2025.5] Biomni: A general-purpose biomedical ai agent [paper]
  132. [arxiv 2025.4] Llm agent swarm for hypothesis-driven drug discovery [paper]
  133. [arxiv 2025.4] Towards a HIPAA Compliant Agentic AI System in Healthcare [paper]
  134. [arxiv 2025.4] Customizing emotional support: How do individuals construct and interact with LLM-powered chatbots [paper]
  135. [arxiv 2025.4] Privacy-Preserving Operating Room Workflow Analysis using Digital Twins [paper]
  136. [arxiv 2025.4] An LLM-Driven Multi-Agent Debate System for Mendelian Diseases [paper]
  137. [arxiv 2025.4] Txgemma: Efficient and agentic llms for therapeutics [paper]
  138. [medrxiv 2025.4] TrialGenie: Empowering Clinical Trial Design with Agentic Intelligence and Real World Data [paper]
  139. [MICCAI 2025] Operating room workflow analysis via reasoning segmentation over digital twins [paper]
  140. [arxiv 2025.3] TAMA: A Human--AI Collaborative Thematic Analysis Framework Using Multi-Agent LLMs for Clinical Interviews [paper]
  141. [arxiv 2025.3] Autonomous Radiotherapy Treatment Planning Using DOLA: A Privacy-Preserving, LLM-Based Optimization Agent [paper]
  142. [arxiv 2025.3] The Application of MATEC (Multi-AI Agent Team Care) Framework in Sepsis Care [paper]
  143. [EMNLP 2025] MDTeamGPT: A Self-Evolving LLM-Based Multi-Agent Framework for Multi-Disciplinary Team Medical Consultation [paper] [GitHub]
  144. [arxiv 2025.3] RAG-KG-IL: A Multi-Agent Hybrid Framework for Reducing Hallucinations and Enhancing LLM Reasoning through RAG and Incremental Knowledge Graph Learning Integration [paper]
  145. [arxiv 2025.3] MAP: Evaluation and Multi-Agent Enhancement of Large Language Models for Inpatient Pathways [paper]
  146. [ICASSP 2025] A Self-Evolving Framework for Multi-Agent Medical Consultation Based on Large Language Models [paper]
  147. [arxiv 2025.3] TxAgent: An AI agent for therapeutic reasoning across a universe of tools [paper]
  148. [arxiv 2025.3] MedAgentsBench: Benchmarking Thinking Models and Agent Frameworks for Complex Medical Reasoning [paper] [project page]
  149. [arxiv 2025.3] Towards conversational ai for disease management [paper]
  150. [arxiv 2025.3] GEMA-Score: Granular Explainable Multi-Agent Score for Radiology Report Evaluation [paper]
  151. [EMNLP 2025 Findings] MIND: Towards Immersive Psychological Healing with Multi-Agent Inner Dialogue [paper]
  152. [arxiv 2025.2] Enhancing hepatopathy clinical trial efficiency: a secure, large language model-powered pre-screening pipeline [paper]
  153. [arxiv 2025.2] RAG-Enhanced Collaborative LLM Agents for Drug Discovery [paper]
  154. [EMNLP 2025 Findings] Agentic Medical Knowledge Graphs Enhance Medical Question Answering: Bridging the Gap Between LLMs and Evolving Medical Knowledge [paper]
  155. [arxiv 2025.2] An LLM-Powered Agent for Physiological Data Analysis: A Case Study on PPG-based Heart Rate Estimation [paper]
  156. [arxiv 2025.2] Regulatory science innovation for generative AI and large language models in health and medicine: a global call for action [paper]
  157. [ACL 2025] Cami: A counselor agent supporting motivational interviewing through state inference and topic exploration [paper]
  158. [ICML 2025] MedRAX: Medical Reasoning Agent for Chest X-ray [paper] [code]
  159. [ICCV 2025] PathFinder: A Multi-Modal Multi-Agent System for Medical Diagnostic Decision-Making Applied to Histopathology [Paper] [project page] [Github]
  160. [arxiv 2025.2] M^3Builder: A Multi-Agent System for Automated Machine Learning in Medical Imaging [paper]
  161. [NEJM AI 2025] MedAgentBench: A Realistic Virtual EHR Environment to Benchmark Medical LLM Agents [paper] [project page]
  162. [arxiv 2025.1] AI Chatbots as Professional Service Agents: Developing a Professional Identity [paper]
  163. [arxiv 2025.1] Exploring the inquiry-diagnosis relationship with advanced patient simulators [paper] [project page]
  164. [ICML 2025] MedXpertQA: Benchmarking Expert-Level Medical Reasoning and Understanding [paper] [project page]
  165. [arxiv 2025.1] AutoCBT: An Autonomous Multi-agent Framework for Cognitive Behavioral Therapy in Psychological Counseling [paper]
  166. [medrxiv 2025.1] Advancing the prediction and understanding of placebo responses in chronic back pain using large language models [paper]
  167. [Nature] Towards conversational diagnostic artificial intelligence [paper]
  168. [Nature Communications 2025] AgentMD: Empowering Language Agents for Risk Prediction with Large-Scale Clinical Tool Learning [paper]
  169. [Intelligent Medicine] Evaluating large language models and agents in healthcare: key challenges in clinical applications [paper]
  170. [npj Digital Medicine] Evaluating large language models as agents in the clinic [paper]
  171. [Nature Medicine 2025] An evaluation framework for clinical use of large language models in patient interaction tasks [paper]
  172. [Nature Communications 2025] An automated framework for assessing how well LLMs cite relevant medical references [paper]
  173. [Nature BME 2025] CRISPR-GPT for agentic automation of gene-editing experiments [paper]
  174. [Nature Methods 2025] GeneAgent: self-verification language agent for gene-set analysis using domain databases [paper]
  175. [npj Digital Medicine] CARE-AD: A Multi-Agent Large Language Model Framework for Alzheimer's Disease Prediction Using Longitudinal Clinical Notes [paper]
  176. [npj Digital Medicine] Vision-language model for report generation and outcome prediction in CT pulmonary angiogram [paper]
  177. [npj Artificial Intelligence] HealthcareAgent: Eliciting the Power of Large Language Models for Medical Consultation [paper]
  178. [Scientific Reports 2025] Democratizing cost-effective, agentic artificial intelligence to multilingual medical summarization through knowledge distillation [paper]
  179. [Scientific Reports 2025] A multi-agent system based on HNC for domain-specific machine translation [paper]
  180. [biorxiv 2025.6] HEAL-KGGen: A Hierarchical Multi-Agent LLM Framework with Knowledge Graph Enhancement for Genetic Biomarker-Based Medical Diagnosis [paper]
  181. [JAMIA 2025] Improving Large Language Model Applications in Biomedicine with Retrieval-Augmented Generation: A Systematic Review, Meta-Analysis, and Clinical Development Guidelines [paper]
  182. [JAMIA Open 2025] Conversational health agents: a personalized large language model-powered agent framework [paper]
  183. [JMIR] The Effectiveness of a Custom AI Chatbot for Type 2 Diabetes Mellitus Health Literacy: Development and Evaluation Study [paper]
  184. [JMIR Aging 2025] The PDC30 Chatbot—Development of a Psychoeducational Resource on Dementia Caregiving Among Family Caregivers: Mixed Methods Acceptability Study [paper]
  185. [JoVE] Evidence-based knowledge synthesis and hypothesis validation: Navigating biomedical knowledge bases via explainable ai and agentic systems [paper]
  186. [arxiv 2024.8] Drugagent: Multi-agent large language model-based reasoning for drug-target interaction prediction [paper]
  187. [Bioinformatics 2025] ESCARGOT: an AI agent leveraging large language models, dynamic graph of thoughts, and biomedical knowledge graphs for enhanced reasoning [paper]
  188. [Healthcare (Basel) 2025] MedScrubCrew: A Medical Multi-Agent Framework for Automating Appointment Scheduling Based on Patient-Provider Profile Resource Matching [paper]
  189. [Clinical Neurophysiology 2025] Agent-guided AI-powered interpretation and reporting of nerve conduction studies and EMG (INSPIRE) [paper]
  190. [Expert Systems with Applications 2025] A two-stage proactive dialogue generator for efficient clinical information collection using large language model [paper]
  191. [Physics in Medicine & Biology 2025] A feasibility study of automating radiotherapy planning with large language model agents [paper]
  192. [JCO 2025] A large language model (LLM)-based multi-agent framework for risk stratification and treatment recommendations in localized prostate cancer (locPCa). [paper]
  193. [ICDH] Voice-based AI Agents: Filling the Economic Gaps in Digital Health Delivery [paper]
  194. [IEEE EMBC 2025] Knowledge-infused LLM-powered conversational health agent: A case study for diabetes patients [paper]
  195. [ICLR 2025] MMed-RAG: Versatile Multimodal RAG System for Medical Vision Language Models [paper]
  196. [ACL 2025] Medical Graph RAG: Evidence-based Medical Large Language Model via Graph Retrieval-Augmented Generation [paper]
  197. [ACL Findings 2025] MAM: Modular Multi-Agent Framework for Multi-Modal Medical Diagnosis via Role-Specialized Collaboration [paper]
  198. [ACL Findings 2025] ASTRID--An Automated and Scalable TRIaD for the Evaluation of RAG-based Clinical Question Answering Systems [paper]
  199. [NAACL 2025] A Layered Debating Multi-Agent System for Similar Disease Diagnosis [paper]
  200. [NAACL 2025] Menti: Bridging medical calculator and llm agent with nested tool calling [paper]
  201. [COLING 2025] Unveiling performance challenges of large language models in low-resource healthcare: A demographic fairness perspective [paper]
  202. [ICMI 2025] An LLM-powered Socially Interactive Agent with Adaptive Facial Expressions for Conversing about Health [paper]
  203. [MICCAI 2025 (Oral)] WSI-Agents: A Collaborative Multi-Agent System for Multi-Modal Whole Slide Image Analysis [Paper] [GitHub]
  204. [MICCAI 2025] Multi-Agent Reasoning for Cardiovascular Imaging Phenotype Analysis [Paper] [GitHub]
  205. [MICCAI 2025] DentEval: Fine-tuning-Free Expert-Aligned Assessment in Dental Education via LLM Agents [Paper] [GitHub]
  206. [MICCAI 2025] CSAP-Assist: Instrument-Agent Dialogue Empowered Vision-Language Models for Collaborative Surgical Action Planning [Paper] [GitHub]
  207. [MICCAI 2025] MedAgentSim: Self-Evolving Multi-Agent Simulations for Realistic Clinical Interactions [Paper] [Github]
  208. [MICCAI 2025 workshop] AURA: A Multi-Modal Medical Agent for Understanding, Reasoning & Annotation [paper] [github]
  209. [ICT4AWE 2025] MentalRAG: Developing an Agentic Framework for Therapeutic Support Systems [paper]
  210. [MLHC 2025] Evaluation of Multi-Agent LLMs in Multidisciplinary Team Decision-Making for Challenging Cancer Cases [paper]
  211. [Journal of imaging informatics in medicine] AgentMRI: A Vison Language Model-Powered AI System for Self-regulating MRI Reconstruction with Multiple Degradations [paper]
  212. [COLM 2025] Can A Society of Generative Agents Simulate Human Behavior and Inform Public Health Policy? A Case Study on Vaccine Hesitancy [paper]
  213. [ACL 2025 Findings] PIORS: Personalized Intelligent Outpatient Reception based on Large Language Model with Multi-Agents Medical Scenario Simulation [paper] [project page]
  214. [Communications Medicine 2025] Simulated patient systems are intelligent when powered by large language model-based AI agents [paper]
  215. [AAMAS 2025] On the limits of agency in agent-based models [paper]
  216. [ACL 2025 Findings] Cod, towards an interpretable medical agent using chain of diagnosis [paper] [Github]
  217. [Advanced Intelligent Systems 2025] Inquire, Interact, and Integrate: A Proactive Agent Collaborative Framework for Zero-Shot Multimodal Medical Reasoning [paper]
  218. [NeurIPS 2025] Clinicallab: Aligning agents for multi-departmental clinical diagnostics in the real world [paper]
  219. [Cell Reports Medicine 2025] Development and Testing of a Novel Large Language Model-Based Clinical Decision Support Systems for Medication Safety in 12 Clinical Specialties [paper]
  220. [PMLR 2025] KG4Diagnosis: A Hierarchical Multi-Agent LLM Framework with Knowledge Graph Enhancement for Medical Diagnosis [paper]
  221. [Nature Machine Intelligence 2025] LLM-based agentic systems in medicine and healthcare [paper]
  222. [ACL 2025 Findings] A Survey of LLM-based Agents in Medicine: How far are we from Baymax? [paper] [Github]
  223. [TechRxiv 2025] The Landscape of Medical Agents: A Survey [paper] [Github]
  224. [TechRxiv 2025] Agentic large-language-model systems in medicine: A systematic review and taxonomy [paper]
  225. [Medicine Advances 2025] Agentic large language models for healthcare: current progress and future opportunities [paper]
  226. [TechRxiv 2025] A Survey of LLM-based Multi-agent Systems in Medicine [paper]
  227. [Cell Reports Medicine 2025] A foundational architecture for AI agents in healthcare [paper]
  228. [Nature Biomedical Engineering 2025] Coordinated AI agents for advancing healthcare [paper]
  229. [Cell Reports Medicine 2025] Next-generation agentic AI for transforming healthcare [paper]
  230. [Information (MDPI) 2025] Large Language Model Agents for Biomedicine: A Comprehensive Review of Methods, Evaluations, Challenges, and Future Directions [paper]
  231. [PLOS ONE 2025] Artificial intelligence agents in healthcare research: A scoping review [paper]
  232. [npj Digital Medicine 2025] Enhancing diagnostic capability with multi-agents conversational large language models [paper] [Github]
  233. [International Journal of Medical Informatics 2025] Applications of artificial intelligence-based conversational agents in healthcare: A systematic umbrella review [paper]
  234. [HAL 2025] Scoping Review of Agentic AI Systems in Healthcare [paper]
  235. [Preprints.org 2025] AI Agents in Modern Healthcare: From Foundation to Pioneer — A Comprehensive Review and Implementation Roadmap for Impact and Integration in Clinical Settings [paper]
  236. [Asian Journal of Medical Principles and Clinical Practice 2025] Multi-Agent AI Systems in Healthcare: A Systematic Review Enhancing Clinical Decision-Making [paper]
  237. [medRxiv 2025] AI agents in clinical medicine: a systematic review [paper]
  238. [Radiology: Artificial Intelligence (RSNA) 2025] Agentic AI in Radiology: Evolution from Large Language Models to Future Clinical Integration [paper]
  239. [Indian Journal of Radiology and Imaging 2025] From chatbots to agentic workflows: ensuring responsible deployment of large language models in radiology [paper]
  240. [Bioengineering (MDPI) 2025] Agentic AI and Large Language Models in Radiology: Opportunities and Hallucination Challenges [paper]
  241. [arxiv 2025.10] Agentic systems in radiology: Design, Applications, Evaluation, and Challenges [paper]
  242. [British Journal of Radiology 2025] Agentic AI in radiology: emerging potential and unresolved challenges [paper]
  243. [Radiography 2025] Agentic systems in radiology: Principles, opportunities, privacy risks, regulation, and sustainability concerns [paper]
  244. [Tomography (MDPI) 2025] The Role of Agentic AI in Musculoskeletal Radiology: A Scoping Review [paper]
  245. [Nurse Education Today 2025] Large language model-driven agents in nursing practice: A scoping review [paper]
  246. [Communications Medicine 2025] Simulated patient systems powered by large language model-based AI agents offer potential for transforming medical education [paper]
  247. [Biocomputing 2025] Using large language models for efficient cancer registry coding in the real hospital setting: A feasibility study [paper]

Year 2024

  1. [arxiv 2024.12] PsyDraw: A Multi-Agent Multimodal System for Mental Health Screening in Left-Behind Children [paper]
  2. [IEEE Big Data] SurgBox: Agent-Driven Operating Room Sandbox with Surgery Copilot [paper] [code]
  3. [Bioinformatics] AI-HOPE: an AI-driven conversational agent for enhanced clinical and genomic data integration in precision medicine research [paper]
  4. [arxiv 2024.10] IMAS: A Comprehensive Agentic Approach to Rural Healthcare Delivery [paper] [project page]
  5. [arxiv 2024.10] KGARevion: An AI Agent for Knowledge-Intensive Biomedical QA [paper] [Github] [Project]
  6. [arxiv 2024.10] Zodiac: A Cardiologist-Level LLM Framework for Multi-Agent Diagnostics [paper]
  7. [arxiv 2024.9] Chatting Up Attachment: Using LLMs to Predict Adult Bonds [paper]
  8. [MLHC 2024] MALADE: Orchestration of LLM-powered Agents with Retrieval Augmented Generation for Pharmacovigilance [paper] [project page]
  9. [arxiv 2024.8] Agentic llm workflows for generating patient-friendly medical reports [paper] [project page]
  10. [ACM UIST 2024] Compeer: A generative conversational agent for proactive peer support [paper]
  11. [arxiv 2024.7] Cactus: Towards psychological counseling conversations using cognitive behavioral theory [paper]
  12. [TMI] Integration of Multi-Source Medical Data for Medical Diagnosis Question Answering [paper]
  13. [ICLR 2025 Oral] Pathgen-1.6m: 1.6 million pathology image-text pairs generation through multi-agent collaboration [paper] [project page]
  14. [arxiv 2024.7] MentalAgora: A Gateway to Advanced Personalized Care in Mental Health through Multi-Agent Debating and Attribute Control [paper]
  15. [arxiv 2024.6] Exploring llm multi-agents for icd coding [paper]
  16. [arxiv 2024.12] Enhancing LLMs for Impression Generation in Radiology Reports through a Multi-Agent System [paper]
  17. [ICML 2024 AI for Science Workshop] TriageAgent: Towards Better Multi-Agents Collaborations for Large Language Model-Based Clinical Triage [paper]
  18. [KDD'24 Workshop] EHRFlow: A Large Language Model-Driven Iterative Multi-Agent Electronic Health Record Data Analysis Workflow [paper]
  19. [arxiv 2024.12] Agents on the Bench: Large Language Model Based Multi-Agent Framework for Trustworthy Digital Justice [paper]
  20. [MLHS 2025] Path-RAG: Knowledge-Guided Key Region Retrieval for Open-ended Pathology Visual Question Answering [paper]
  21. [NeurIPS 2024] MEDIQ: Question-Asking LLMs and a Benchmark for Medical Information-Seeking [paper] [project page]
  22. [arxiv 2024.6] CliBench: A Multifaceted and Multigranular Evaluation of Clinical Diagnosis with LLMs [paper]
  23. [arxiv 2024.5] AgentClinic: a multimodal agent benchmark to evaluate AI in simulated clinical environments [paper]
  24. [AAAI 2025 workshop AI4Research] Drugagent: Automating ai-aided drug discovery programming through llm multi-agent collaboration [paper]
  25. [arxiv 2024.5] Agent Hospital: A Simulacrum of Hospital with Evolvable Medical Agents [paper]
  26. [EMNLP 2024] Ehragent: Code empowers large language models for few-shot complex tabular reasoning on electronic health records [paper]
  27. [NeurIPS 2024 Oral] Mdagents: An adaptive collaboration of llms for medical decision-making [paper] [project page]
  28. [arxiv 2024.3] Llms-based few-shot disease predictions using ehr: A novel approach combining predictive agent reasoning and critical agent instruction [paper]
  29. [npj Digital Medicine] PRISM: Patient Records Interpretation for Semantic Clinical Trial Matching using Large Language Models [paper]
  30. [arxiv 2024.1] A general-purpose AI avatar in healthcare [paper]
  31. [The Lancet Digital Health] A future role for health applications of large language models depends on regulators enforcing safety standards [paper]
  32. [npj Digital Medicine] Autonomous medical evaluation for guideline adherence of large language models [paper]
  33. [Diagn Interv Radiol 2024] Large language models in radiology: fundamentals, applications, ethical considerations, risks, and future directions [paper]
  34. [PACIFIC SYMPOSIUM ON BIOCOMPUTING 2024] A conversational agent for early detection of neurotoxic effects of medications through automated intensive observation [paper]
  35. [JAMIA Open 2024] Conversational health agents: A personalized llm-powered agent framework [paper] [project page]
  36. [JMIR 2024] Mitigating cognitive biases in clinical decision-making through multi-agent conversations using large language models: simulation study [paper]
  37. [JMIR 2024] A language model--powered simulated patient with automated feedback for history taking: Prospective study [paper]
  38. [IEEE SoftCOM 2024] A multi-agent architecture for privacy-preserving natural language interaction with FHIR-based electronic health records [paper]
  39. [IEEE ISDFS 2024] Llm-based framework for administrative task automation in healthcare [paper]
  40. [IEEE Access 2024] Knowledge-Routed Automatic Diagnosis With Heterogeneous Patient-Oriented Graph [paper]
  41. [EMNLP Findings 2024] MMedAgent: Learning to Use Medical Tools with Multi-modal Agent [paper]
  42. [EMNLP 2024] RULE: Reliable Multimodal RAG for Factuality in Medical Vision Language Models [paper] [Github]
  43. [ACL Findings 2024] Benchmarking large language models on communicative medical coaching: a dataset and a novel system [paper]
  44. [ACL Findings 2024] Medagents: Large language models as collaborators for zero-shot medical reasoning [paper]
  45. [AAAI 2024] PathAsst: A Generative Foundation AI Assistant towards Artificial General Intelligence of Pathology [paper] [Github]
  46. [CHI 2024] Understanding the impact of long-term memory on self-disclosure with large language model-driven chatbots for public health intervention [paper]
  47. [CHI EA 2024] Conversational AI in health: Design considerations from a Wizard-of-Oz dermatology case study with users, clinicians and a medical LLM [paper]
  48. [ACM IMWUT 2024] Talk2Care: An LLM-based Voice Assistant for Communication between Healthcare Providers and Older Adults [paper]
  49. [ArabicNLP 2024] Synthetic arabic medical dialogues using advanced multi-agent llm techniques [paper]
  50. [ECCV Workshop 2024] Medco: Medical education copilots based on a multi-agent framework [paper]
  51. [Healthcare Information 2024] A Medical Consultation System for Geriatric Disease Based on Multi-agent Architecture and Knowledge Graph [paper]
  52. [Cell 2024] Empowering biomedical discovery with AI agents [paper] [Github]

Year 2023

  1. [NeurIPS workshop 2023] Are we going mad? benchmarking multi-agent debate between language models for medical q&a [paper]
  2. [arxiv 2023.1] Talk2Care: Facilitating asynchronous patient-provider communication with large-language-model [paper]
  3. [AMIA Annual Symposium Proceedings] Understanding the benefits and challenges of using large language model-based conversational agents for mental well-being support [paper]
  4. [Clinical NLP 2023] DERA: enhancing large language model completions with dialog-enabled resolving agents [paper] [dataset]
  5. [JMIR] The ChatGPT (generative artificial intelligence) revolution has made artificial intelligence approachable for medical professionals [paper]
  6. [JMIR] Automated monitoring of adherence to evidenced-based clinical guideline recommendations: design and implementation study [paper]
  7. [JMIR Med Educ 2023] Using ChatGPT for clinical practice and medical education: cross-sectional survey of medical students’ and physicians’ perceptions [paper]
  8. [CHI 2023] Assertiveness-based agent communication for a personalized medicine on medical imaging diagnosis [paper]

Papers by Category


1. Doctor-facing Agents

1.1 Multi-Modal Clinical Agents

(Agents designed to process and reason over multiple data types like images, text, and structured data)

TitleVenueDatePaper LinkProject Page
MIRA: Medical Image Reflection for Agentic DiagnosisarXiv2026.08PaperNot Available
Inference-Time Agentic Decision Rules Beat Longer Evolving Search for Multi-Image Medical ReasoningCVPR Workshop2026.07PaperNot Available
Understanding From Human Perspective: A Multi-agent System for Interactive Egocentric Medical Image SegmentationarXiv2026.07PaperStar
GitHub
MedRLM: Recursive Multimodal Health Intelligence for Long-Context Clinical Reasoning, Sensor-Guided Screening, Evidence-Grounded Decision Support, and Community-to-Tertiary Referral OptimizationarXiv2026.06PaperNot Available
XMedFusion: A Knowledge-Guided Multimodal Perception and Reasoning Framework for Autonomous Medical SystemsarXiv2026.06PaperNot Available
ArogyaSutra: A Multi-Agent Framework for Multimodal Medical Reasoning in Indic LanguagesIJCAI2026.06PaperNot Available
Towards Conversational Medical AI with Eyes, Ears and a VoicearXiv2026.05PaperNot Available
VERITAS: A Multi-Agent Co-Scientist for Verifiable Image-Derived Hypothesis TestingarXiv2026.04PaperStar
GitHub
Camyla: Scaling Autonomous Research in Medical Image SegmentationarXiv2026.04PaperProject
MedVR: Annotation-Free Medical Visual Reasoning via Agentic Reinforcement LearningICLR2026.04PaperNot Available
MedOpenClaw: Auditable Medical Imaging Agents Reasoning over Uncurated Full StudiesarXiv2026.03PaperStar
GitHub
Project
Cerebra: A Multidisciplinary AI Board for Multimodal Dementia Characterization and Risk AssessmentarXiv2026.03PaperNot Available
Shifting Adaptation from Weight Space to Memory Space: A Memory-Augmented Agent for Medical Image SegmentationarXiv2026.03PaperNot Available
Evolving Medical Imaging Agents via Experience-driven Self-skill DiscoveryarXiv2026.03PaperNot Available
Towards a Medical AI ScientistarXiv2026.03PaperProject
Meissa: Multi-modal Medical Agentic IntelligencearXiv2026.03PaperStar
GitHub
CARE: Towards Clinical Accountability in Multi-Modal Medical ReasoningICLR2026.03PaperProject
3DMedAgent: Unified Perception-to-Understanding for 3D Medical AnalysisarXiv2026.02PaperNot Available
CoMMa: Contribution-Aware Medical Multi-Agents From A Game-Theoretic PerspectivearXiv2026.02PaperNot Available
MedXIAOHE: A Comprehensive Recipe for Building Medical MLLMsarXiv2026.02PaperNot Available
Picking the Right Specialist: Attentive Neural Process-based Selection of Task-Specialized ModelsarXiv2026.02PaperNot Available
Human-Guided Agentic AI for Multimodal Clinical PredictionICHI2026.02PaperNot Available
MedSAM-Agent: Empowering Interactive Medical Image Segmentation with Multi-turn Agentic RLarXiv2026.02PaperStar
GitHub
IBISAgent: Reinforcing Pixel-Level Visual Reasoning in MLLMsarXiv2026.01PaperNot Available
MedEyes: Learning Dynamic Visual Focus for Medical Progressive DiagnosisarXiv2025.11PaperStar
GitHub
MedSAM3: Delving into Segment Anything with Medical ConceptsarXiv2025.11PaperStar
GitHub
AURA: A Multi-modal Medical Agent for Understanding, Reasoning & AnnotationMICCAI workshop2025.07PaperStar
GitHub
MedAgent-Pro: Towards Evidence-based Multi-modal Medical Diagnosis via Reasoning Agentic WorkflowarXiv2025.03PaperStar
GitHub
M^3Builder: A Multi-Agent System for Automated Machine Learning in Medical ImagingarXiv2025.02PaperStar
GitHub
MAM: Modular Multi-Agent Framework for Multi-Modal Medical Diagnosis via Role-Specialized CollaborationACL2025PaperStar
GitHub
MedAgentSim: Self-Evolving Multi-Agent Simulations for Realistic Clinical InteractionsMICCAI2025PaperStar
GitHub
MDAgents: An Adaptive Collaboration of LLMs for Medical Decision-MakingNeurIPS (Oral)2024PaperStar
GitHub
MMedAgent: Learning to Use Medical Tools with Multi-modal AgentEMNLP Findings2024PaperStar
GitHub

1.2 Radiology Agents (CT, X-ray, MRI, etc.)

TitleVenueDatePaper LinkProject Page
Policy-Driven CT-Agent: Modeling Phase-Aware Diagnostic Control for Clinically Consistent CT ReasoningarXiv2026.07PaperNot Available
CogRad: A Cognitively-Inspired Multi-Agent Framework for Radiology Report GenerationarXiv2026.07PaperNot Available
A multi-agent system for spine MRI report generation from multi-sequence imagingarXiv2026.06PaperNot Available
ABRA: Agent Benchmark for Radiology ApplicationsarXiv2026.05PaperNot Available
DeepTumorVQA: A Hierarchical 3D CT Benchmark for Stage-Wise Evaluation of Medical VLMs and Tool-Augmented AgentsarXiv2026.05PaperNot Available
GAZE: Grounded Agentic Zero-shot Evaluation with Viewer-Level Tools and Literature Retrieval on Rare Brain MRIarXiv2026.05PaperNot Available
Agentic Large Language Models for Training-Free Neuro-Radiological Image AnalysisarXiv2026.04PaperNot Available
MARCH: Multi-Agent Radiology Clinical Hierarchy for CT Report GenerationACL2026.04PaperNot Available
RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomographyarXiv2026.04PaperNot Available
Evo-MedAgent: Beyond One-Shot Diagnosis with Agents That Remember, Reflect, and ImprovearXiv2026.04PaperNot Available
XrayClaw: Cooperative-Competitive Multi-Agent Alignment for Trustworthy Chest X-ray DiagnosisarXiv2026.04PaperNot Available
EviAgent: Evidence-Driven Agent for Radiology Report GenerationarXiv2026.03PaperNot Available
Agentic Automation of BT-RADS Scoring: End-to-End Multi-Agent System for Standardized Brain Tumor Follow-up AssessmentarXiv2026.03PaperNot Available
DUCX: Decomposing Unfairness in Tool-Using Chest X-ray AgentsarXiv2026.03PaperNot Available
Can Agents Distinguish Visually Hard-to-Separate Diseases in a Zero-Shot Setting?arXiv2026.02PaperStar
GitHub
Which Tool Response Should I Trust? Tool-Expertise-Aware CXR Agent with Multimodal Agentic LearningarXiv2026.02PaperNot Available
Perfusion Imaging and Single Material Reconstruction in Polychromatic Photon Counting CTarXiv2026.02PaperStar
GitHub
Route, Retrieve, Reflect, Repair: Self-Improving Agentic Framework for Visual DetectionarXiv2026.01PaperStar
GitHub
Explainable Agentic AI Framework for Acute Ischemic Stroke Imaging DecisionsarXiv2026.01PaperNot Available
LungNoduleAgent: A Collaborative Multi-Agent System for Precision Diagnosis of Lung NodulesAAAI2026.1PaperStar
GitHub
Bidirectional human-AI collaboration in brain tumour assessments improves both expert human and AI agent performancearXiv2025.12PaperNot Available
INFORM-CT: INtegrating LLMs and VLMs FOR Incidental Findings Management in Abdominal CTarXiv2025.12PaperNot Available
Radiologist Copilot: Agentic AI Assistant for Holistic Radiology Reporting with Quality ControlarXiv2025.12PaperNot Available
A Multi-Agent System for Complex Reasoning in Radiology Visual Question AnsweringarXiv2025.08PaperNot Available
AT-CXR: Uncertainty-Aware Agentic Triage for Chest X-raysarXiv2025.08PaperStar
GitHub
PASS: Probabilistic Agentic Supernet Sampling for Interpretable and Adaptive Chest X-Ray ReasoningarXiv2025.08PaperStar
GitHub
RadFabric: Agentic AI System with Reasoning Capability for RadiologyarXiv2025.06PaperProject
A Multimodal Multi-Agent Framework for Radiology Report GenerationarXiv2025.05PaperNot Available
CT-Agent: A Multimodal-LLM Agent for 3D CT Radiology Question AnsweringarXiv2025.05PaperNot Available
MedRAX: Medical reasoning agent for chest x-rayICML2025.02PaperStar
GitHub
Vision-language model for report generation and outcome prediction in CT pulmonary angiogramnpj Digital Medicine2025PaperStar
GitHub
AgentMRI: A Vison Language Model-Powered AI System for Self-regulating MRI Reconstruction with Multiple DegradationsJournal of imaging informatics in medicine2025PaperNot Available
Enhancing LLMs for Impression Generation in Radiology Reports through a Multi-Agent SystemarXiv2024.12PaperNot Available

1.3 Pathology Agents

TitleVenueDatePaper LinkProject Page
Trust but Verify:Evidence-Linked Multi-Agent Clinical Information Extraction in PathologyarXiv2026.07PaperNot Available
Democratizing and accelerating AI-driven pathology research through agentic intelligencearXiv2026.06PaperNot Available
Prompt, Plan, Extract: Zero-Shot Agentic LLMs Workflows for Lung Pathology Extraction from Clinical NarrativesarXiv2026.06PaperNot Available
A Multi-modal Agentic Co-pilot for Evidence Grounded Computational PathologyarXiv2026.06PaperNot Available
Computational Pathology in the Era of Emerging Foundation and Agentic AI -- International Expert PerspectivesarXiv2026.03PaperNot Available
LAMMI-Pathology: A Tool-Centric Bottom-Up LVLM-Agent Framework for Molecularly Informed Medical IntelligencearXiv2026.02PaperNot Available
SurvAgent: Hierarchical CoT-Enhanced Case Banking and Dichotomy-Based Multi-Agent System for Multimodal Survival PredictionarXiv2025.11PaperNot Available
GMAT: Grounded Multi-Agent Clinical Description Generation for Text Encoder in Vision-Language MILarXiv2025.08PaperNot Available
Patho-AgenticRAG: Towards Multimodal Agentic Retrieval-Augmented Generation for Pathology VLMsarXiv2025.08PaperStar
GitHub
Evidence-based diagnostic reasoning with multi-agent copilot for human pathologyarXiv2025.06PaperNot Available
CPathAgent: An Agent-based Foundation Model for Interpretable High-Resolution Pathology Image AnalysisNeurIPS2025.05PaperNot Available
PathFinder: A Multi-Modal Multi-Agent System for Medical Diagnostic Decision-Making Applied to HistopathologyICCV2025.02Paperproject Star
GitHub
WSI-Agents: A Collaborative Multi-Agent System for Multi-Modal Whole Slide Image AnalysisMICCAI (Oral)2025PaperStar
GitHub
Path-RAG: Knowledge-Guided Key Region Retrieval for Open-ended Pathology Visual Question AnsweringMLHS2025PaperStar
GitHub
Pathgen-1.6m: 1.6 million pathology image-text pairs generation through multi-agent collaborationICLR (Oral)2024PaperStar
GitHub
PathAsst: A Generative Foundation AI Assistant towards Artificial General Intelligence of PathologyAAAI2024PaperStar
GitHub

1.4 Cardiovascular Imaging

TitleVenueDatePaper LinkProject Page
Diagnosing as Cardiologists Do: ECG Agents with Doctor-Grounded Priors for Clinical Reasoning Across Diseases and PopulationsarXiv2026.08PaperNot Available
Cardiologent: Multi-Agent Clinical Decision Support for Patient-Level Arrhythmia Assessment, Urgency, and ManagementarXiv2026.07PaperNot Available
ECG Foundation Models and Medical LLMs for Agentic Cardiovascular Intelligence at the Edge: A Review and OutlookarXiv2026.04PaperNot Available
Multi-Agent Reasoning for Cardiovascular Imaging Phenotype AnalysisMICCAI2025.07PaperStar
GitHub

1.5 Sonography / Ultrasound

TitleVenueDatePaper LinkProject Page
Auditable agentic AI for evidence-grounded thyroid ultrasound diagnosis and reportingarXiv2026.08PaperNot Available
Echo-α: Large Agentic Multimodal Reasoning Model for Ultrasound InterpretationarXiv2026.04PaperNot Available
Anatomical Prior-Driven Framework for Autonomous Robotic Cardiac Ultrasound Standard View AcquisitionICRA2026.03PaperNot Available
Intelligent Virtual Sonographer (IVS): Enhancing Physician-Robot-Patient CommunicationarXiv2025.07PaperStar
GitHub

1.6 Radiotherapy

TitleVenueDatePaper LinkProject Page
Autonomous Radiotherapy Treatment Planning Using DOLA: A Privacy-Preserving, LLM-Based Optimization AgentarXiv2025.03PaperNot Available
A feasibility study of automating radiotherapy planning with large language model agentsPhysics in Medicine & Biology2025PaperNot Available

1.7 Dermatology

TitleVenueDatePaper LinkProject Page
DermAgent: A Self-Reflective Agentic System for Dermatological Image Analysis with Multi-Tool Reasoning and Traceable Decision-MakingMICCAI2026.05PaperNot Available
Conversational AI in health: Design considerations from a Wizard-of-Oz dermatology case study with users, clinicians and a medical LLMCHI 'EA2024PaperNot Available

1.8 Dental Agents

TitleVenueDatePaper LinkProject Page
OPGAgent: An Agent for Auditable Dental Panoramic X-ray InterpretationarXiv2026.03PaperNot Available
DentEval: Fine-tuning-Free Expert-Aligned Assessment in Dental Education via LLM AgentsMICCAI2025PaperStar
GitHub

1.9 Genomics & Biomarker Agents

TitleVenueDatePaper LinkProject Page
CPAgents: Agentic Composite Phenotype Generation for Cardiac Disease AssociationMICCAI2026.06PaperNot Available
DeepBD: A Grounded Agentic Workflow for Variant Prioritization and Diagnosis of Genetic Birth DefectsarXiv2026.06PaperNot Available
Autonomous Agent-Orchestrated Digital Twins (AADT): State Synchronization in Rare Genetic DisordersarXiv2026.03PaperNot Available
ProtRLSearch: A Multi-Round Multimodal Protein Search Agent with LLMs Trained via RLarXiv2026.03PaperNot Available
Geneagent: self-verification language agent for gene-set analysis using domain databasesNature Methods2025PaperStar
GitHub
CRISPR-GPT for agentic automation of gene-editing experimentsNature BME2025PaperStar
GitHub
HEAL-KGGen: A Hierarchical Multi-Agent LLM Framework for Genetic Biomarker-Based Medical Diagnosisbiorxiv2025PaperStar
GitHub
AI-HOPE: An AI-Driven conversational agent for enhanced clinical and genomic data integrationBioinformatics2024.12PaperStar
GitHub
dna-claude-analysis: AI-powered personal genome analysis agent using ClaudeGitHub2025Not AvailableStar
GitHub

1.10 EHR & Clinical Note Agents

TitleVenueDatePaper LinkProject Page
A Multi-Agent System for Autonomous, Fine-Tuning-Free Clinical Symptom Detection: Development and Validation StudyarXiv2026.07PaperNot Available
Multi-Large Language Model Orchestrated Severity Assessment of Clinical Records (MOSAIC)arXiv2026.07PaperNot Available
Configurable Clinical Information Extraction with Agentic RAG: What Works, What Breaks, and WhyarXiv2026.06PaperNot Available
COTCAgent: Preventive Consultation via Probabilistic Chain-of-Thought CompletionarXiv2026.05PaperNot Available
Reinforcement Learning for Tool-Calling Agents in Fast Healthcare Interoperability Resources (FHIR)arXiv2026.05PaperNot Available
Generating synthetic electronic health record data using agent-based models to evaluate machine learning robustness under mass casualty incidentsCHIL2026.05PaperNot Available
CuraView: A Multi-Agent Framework for Medical Hallucination Detection with GraphRAG-Enhanced Knowledge VerificationarXiv2026.05PaperNot Available
PhysicianBench: Evaluating LLM Agents in Real-World EHR EnvironmentsarXiv2026.05PaperNot Available
Clinically Interpretable Sepsis Early Warning via LLM-Guided Simulation of Temporal Physiological DynamicsarXiv2026.04PaperNot Available
BLUEmed: Retrieval-Augmented Multi-Agent Debate for Clinical Error DetectionIEEE ICHI2026.04PaperNot Available
Beyond the Individual: Virtualizing Multi-Disciplinary Reasoning for Clinical Intake via Collaborative AgentsACL'26 Findings2026.04PaperStar
GitHub
Symphony for Medical Coding: A Next-Generation Agentic System for Scalable and Explainable Medical CodingarXiv2026.03PaperNot Available
Can LLM Agents Generate Real-World Evidence? Evaluating Observational Studies in Medical DatabasesarXiv2026.03PaperStar
GitHub
From Physician Expertise to Clinical Agents: Preserving, Standardizing, and Scaling Physicians' Medical ExpertisearXiv2026.03PaperNot Available
Empowering Locally Deployable Medical Agent via State Enhanced Logical Skills for FHIR-based Clinical TasksarXiv2026.03PaperNot Available
When OpenClaw Meets Hospital: Toward an Agentic Operating System for Dynamic Clinical WorkflowsarXiv2026.03PaperNot Available
TRACE: Temporal Reasoning via Agentic Context Evolution for Streaming EHRsarXiv2026.02PaperNot Available
AgentEHR: Advancing Autonomous Clinical Decision-Making via Retrospective SummarizationarXiv2026.01PaperNot Available
ExperienceWeaver: Optimizing Small-sample Experience Learning for Clinical Text ImprovementarXiv2026.02PaperNot Available
Hybrid-Code: A Privacy-Preserving, Redundant Multi-Agent Framework for Reliable Local Clinical CodingarXiv2025.12PaperNot Available
HARMON-E: Hierarchical Agentic Reasoning for Multimodal Oncology Notes to Extract Structured DataarXiv2025.12PaperNot Available
ClinNoteAgents: An LLM Multi-Agent System for Predicting and Interpreting Heart Failure 30-Day Readmission from Clinical NotesarXiv2025.12PaperNot Available
MedDCR: Learning to Design Agentic Workflows for Medical CodingarXiv2025.11PaperNot Available
OEMA: Ontology-Enhanced Multi-Agent Collaboration Framework for Zero-Shot Clinical Named Entity RecognitionarXiv2025.11PaperNot Available
Grounded by Experience: Generative Healthcare Prediction Augmented with Hierarchical Agentic RetrievalarXiv2025.11PaperNot Available
Traj-CoA: Patient Trajectory Modeling via Chain-of-Agents for Lung Cancer Risk PredictionNeurIPS'25 Workshop2025.10PaperNot Available
Automated Clinical Problem Detection from SOAP Notes using a Collaborative Multi-Agent LLM ArchitecturearXiv2025.08PaperNot Available
SNOW: Agent-Based Feature Generation from Clinical Notes for Outcome PredictionarXiv2025.08PaperProject
Trustworthy Agents for Electronic Health Records through Confidence EstimationarXiv2025.8PaperStar
GitHub
Infherno: End-to-end agent-based FHIR resource synthesis from free-form clinical notesarXiv2025.07PaperStar
GitHub
From EHRs to Patient Pathways: Scalable Modeling of Longitudinal Health Trajectories with LLMsarXiv2025.6PaperNot Available
CARE-AD: a multi-agent large language model framework for Alzheimer’s disease predictionnpj Digital Medicine2025PaperStar
GitHub
Colacare: Enhancing electronic health record modeling through large language model-driven multi-agent collaborationarXiv2024.10Paper[project]
EHRFlow: A Large Language Model-Driven Iterative Multi-Agent Electronic Health Record Data Analysis WorkflowKDD'24 Workshop2024.06PaperStar
GitHub
A multi-agent architecture for privacy-preserving natural language interaction with FHIR-based electronic health recordsIEEE SoftCOM2024PaperNot Available

1.11 Surgical Agents

TitleVenueDatePaper LinkProject Page
Agent-Guided Relational Concept Discovery: Toward Interpretable Surgical Margin AssessmentMICCAI2026.07PaperNot Available
CSAP-Assist: Instrument-Agent Dialogue Empowered Vision-Language Models for Collaborative Surgical Action PlanningMICCAI2025PaperStar
GitHub
Privacy-Preserving Operating Room Workflow Analysis using Digital TwinsarXiv2025.4PaperNot Available

1.12 Education Agents

Related free course: BioDockify Learn - AI in Healthcare: Diagnosis to Drug Discovery - 24 free AI-narrated video lessons covering explainable AI in clinical settings (SHAP, GradCAM, GEMEX), EHR modeling, wearables, and clinical-judgment training with automation-bias scenarios.

TitleVenueDatePaper LinkProject Page
MedEasy: Designing AI Standardized Patients for Clinical Consultation TrainingarXiv2026.06PaperNot Available
Rethinking Patient Education as Multi-turn Multi-modal InteractionarXiv2026.04PaperNot Available
Persona-Based Requirements Engineering for Explainable Multi-Agent Educational Systems: A Scenario Simulator for Clinical Reasoning TrainingCSTE2026.04PaperNot Available
Dialogue to Question Generation for Evidence-based Medical Guideline Agent DevelopmentML4H2026.03PaperNot Available
An Agentic AI Framework for Training General Practitioner Student SkillsarXiv2025.12PaperNot Available
MedTutor-R1: Socratic Personalized Medical Teaching with Multi-Agent SimulationarXiv2025.12PaperStar
GitHub
Exploring Community-Powered Conversational Agent for Health Knowledge AcquisitionarXiv2025.12PaperNot Available

1.13 Reasoning & Multi Agent Techniques

TitleVenueDatePaper LinkProject Page
MARC v1: An Open-Source Multi-Agent Framework for Clinical AI Reasoning and CoordinationarXiv2026.08PaperStar
GitHub
Social Chain of Thought: A Multi-Agent Architecture Grounded in Medical Differential Diagnosis MethodologyarXiv2026.08PaperNot Available
Harnessing Disagreement: Detecting Correlated Agreement Blindness in Multi-Agent TriagearXiv2026.07PaperNot Available
MedCalc-Pro: Solving Complex Medical Calculations with LLM AgentsarXiv2026.07PaperNot Available
DEEPMED Search: An Open-Source Agentic Platform for Medical Deep Research with Introspective VerificationIJCAI2026.06PaperNot Available
MedGuards: Multi-Agent System for Reliable Medical Error Detection and CorrectionarXiv2026.06PaperStar
GitHub
Evo-RAD: Navigating Rare Retinal Disease Diagnosis via Self-Evolving Agentic RetrievalMICCAI2026.06PaperStar
GitHub
Agentic AI-based Framework for Mitigating Premature Diagnostic Handoff and Silent Hallucination in Healthcare ApplicationsarXiv2026.06PaperNot Available
Teaching agentic AI to learn expert reasoning for rare disease diagnosisarXiv2026.06PaperNot Available
Let LLMs Judge Each Other: Multi-Agent Peer-Reviewed Reasoning for Medical Question AnsweringarXiv2026.06PaperNot Available
Trust but Verify: Mitigating Medical Hallucinations via Post-Hoc Adversarial Auditing and Multi-Agent Feedback LoopsarXiv2026.06PaperNot Available
MedLatentDx: Latent Multi-Agent Communication for Cross-Hospital Rare-Disease DiagnosisarXiv2026.06PaperNot Available
Experience Makes Skillful: Enabling Generalizable Medical Agent Reasoning via Self-Evolving Skill MemoryarXiv2026.06PaperNot Available
Baichuan-M4: A Clinical-Grade Medical Agent System for Continuous CarearXiv2026.06PaperNot Available
D2MDT: Department-aware Multidisciplinary Team Consultation with Deliberation for Efficient Clinical PredictionarXiv2026.06PaperNot Available
MeDxAgent: Multi-Agent Consultation for Interactive Medical DiagnosisarXiv2026.06PaperNot Available
SEMA-RAG: A Self-Evolving Multi-Agent Retrieval-Augmented Generation Framework for Medical ReasoningarXiv2026.05PaperNot Available
MedExAgent: Training LLM Agents to Ask, Examine, and Diagnose in Noisy Clinical EnvironmentsarXiv2026.05PaperNot Available
Thinking Like a Clinician: A Cognitive AI Agent for Clinical Diagnosis via Panoramic Profiling and Adversarial DebatearXiv2026.04PaperNot Available
Neuro-Symbolic Resolution of Recommendation Conflicts in Multimorbidity Clinical GuidelinesAAAI Bridge2026.04PaperNot Available
DeepER-Med: Advancing Deep Evidence-Based Research in Medicine Through Agentic AIarXiv2026.04PaperNot Available
QuarkMedSearch: A Long-Horizon Deep Search Agent for Exploring Medical IntelligencearXiv2026.04PaperNot Available
Dialectic-Med: Mitigating Diagnostic Hallucinations via Counterfactual Adversarial Multi-Agent DebateACL2026.04PaperNot Available
Joint Optimization of Reasoning and Dual-Memory for Self-Learning Diagnostic AgentarXiv2026.04PaperNot Available
CARE: Privacy-Compliant Agentic Reasoning with Evidence DiscordancearXiv2026.04PaperNot Available
Improving Clinical Diagnosis with Counterfactual Multi-Agent ReasoningarXiv2026.03PaperNot Available
MediHive: A Decentralized Agent Collective for Medical ReasoningIEEE ICHI2026.03PaperNot Available
ClinicalAgents: Multi-Agent Orchestration for Clinical Decision Making with Dual-MemoryarXiv2026.03PaperNot Available
Multi-Agent Reasoning with Consistency Verification Improves Uncertainty Calibration in Medical MCQAarXiv2026.03PaperNot Available
CarePilot: A Multi-Agent Framework for Long-Horizon Computer Task Automation in HealthcareCVPR Findings2026.03PaperNot Available
Unified-MAS: Universally Generating Domain-Specific Nodes for Empowering Automatic Multi-Agent SystemsarXiv2026.03PaperStar
GitHub
TheraAgent: Multi-Agent Framework with Self-Evolving Memory for PET TheranosticsarXiv2026.03PaperNot Available
OpenHospital: A Thing-in-itself Arena for Evolving and Benchmarking LLM-based Collective IntelligencearXiv2026.03PaperNot Available
MedScope: Incentivizing "Think with Videos" for Clinical Reasoning via Coarse-to-Fine Tool CallingarXiv2026.02PaperNot Available
ATPO: Adaptive Tree Policy Optimization for Multi-Turn Medical DialogueICLR2026.03PaperNot Available
MedCoRAG: Interpretable Hepatology Diagnosis via Hybrid Evidence Retrieval and Multispecialty ConsensusarXiv2026.03PaperNot Available
MedCollab: Causal-Driven Multi-Agent Collaboration for Full-Cycle Clinical DiagnosisarXiv2026.03PaperNot Available
From Conflict to Consensus: Boosting Medical Reasoning via Multi-Round Agentic RAGarXiv2026.03PaperStar
GitHub
TARSE: Test-Time Adaptation via Retrieval of Skills and Experience for Reasoning AgentsarXiv2026.03PaperNot Available
A Multi-Agent Framework for Interpreting Multivariate Physiological Time SeriesarXiv2026.03PaperNot Available
Do Mixed-Vendor Multi-Agent LLMs Improve Clinical Diagnosis?EACL Workshop2026.03PaperNot Available
MedClarify: An Information-Seeking AI Agent for Medical DiagnosisarXiv2026.02PaperNot Available
MedCoG: Maximizing LLM Inference Density in Medical Reasoning via Meta-Cognitive RegulationarXiv2026.02PaperNot Available
Closing Reasoning Gaps in Clinical Agents with Differential Reasoning LearningarXiv2026.02PaperNot Available
A Multi-Agent Framework for Medical AI: Leveraging GPT, LLaMA, and DeepSeek R1arXiv2026.02PaperNot Available
Pruning Minimal Reasoning Graphs for Efficient Retrieval-Augmented GenerationarXiv2026.02PaperNot Available
RE-MCDF: Closed-Loop Multi-Expert LLM Reasoning for Knowledge-Grounded Clinical DiagnosisarXiv2026.02PaperNot Available
Agentic Reasoning for Large Language ModelsarXiv2026.01PaperStar
GitHub
EvoClinician: A Self-Evolving Agent for Multi-Turn Medical DiagnosisarXiv2026.01PaperStar
GitHub
Scaling Medical Reasoning Verification via Tool-Integrated Reinforcement LearningarXiv2026.01PaperNot Available
DEEPMED: Building a Medical DeepResearch Agent via Multi-hop Med-Search DataarXiv2026.01PaperNot Available
Multi-Aspect Knowledge-Enhanced Medical Vision-Language Pretraining with Multi-Agent Data GenerationarXiv2025.12PaperNot Available
Incentivizing Tool-augmented Thinking with Images for Medical Image AnalysisarXiv2025.12PaperNot Available
AdaSearch: Balancing Parametric Knowledge and Search in Large Language Models via Reinforcement LearningarXiv2025.12PaperStar
Github
Multi-Agent Medical Decision Consensus Matrix System: An Intelligent Collaborative Framework for Oncology MDT ConsultationsarXiv2025.12PaperNot Available
Multi-Agent Intelligence for Multidisciplinary Decision-Making in Gastrointestinal OncologyarXiv2025.12PaperNot Available
DART: Leveraging Multi-Agent Disagreement for Tool Recruitment in Multimodal ReasoningarXiv2025.12PaperStar
Github
MCP-AI: Protocol-Driven Intelligence Framework for Autonomous Reasoning in HealthcarearXiv2025.12PaperNot Available
Many-to-One Adversarial Consensus: Exposing Multi-Agent Collusion Risks in AI-Based HealthcarearXiv2025.12PaperNot Available
Thucy: An LLM-based Multi-Agent System for Claim Verification across Relational DatabasesAAAI Workshop2025.12PaperNot Available
UCAgents: Unidirectional Convergence for Visual Evidence Anchored Multi-Agent Medical Decision-MakingarXiv2025.12PaperStar
GitHub
KOM: A Multi-Agent Artificial Intelligence System for Precision Management of Knee Osteoarthritis (KOA)arXiv2025.11PaperNot Available
KRAL: Knowledge and Reasoning Augmented Learning for LLM-assisted Clinical Antimicrobial TherapyarXiv2025.11PaperNot Available
MedResearcher-R1: Expert-Level Medical Deep Researcher via A Knowledge-Informed Trajectory Synthesis FrameworkarXiv2025.8PaperStar
GitHub
ConfAgents: A Conformal-Guided Multi-Agent Framework for Cost-Efficient Medical DiagnosisarXiv2025.8PaperStar
GitHub
Tree-of-Reasoning: Towards Complex Medical Diagnosis via Multi-Agent Reasoning with Evidence TreearXiv2025.8PaperStar
GitHub
End-to-End Agentic RAG System Training for Traceable Diagnostic ReasoningarXiv2025.8PaperStar
GitHub
A Multi-Agent Approach to Neurological Clinical ReasoningarXiv2025.8PaperNot Available
KERAP: A knowledge-enhanced reasoning approach for accurate zero-shot diagnosis predictionarXiv2025.7PaperStar
GitHub
MMedAgent-RL: Optimizing Multi-Agent Collaboration for Multimodal Medical ReasoningarXiv2025.06PaperNot Available
An agentic system for rare disease diagnosis with traceable reasoningarXiv2025.6Paper[demo]
MedOrch: Medical Diagnosis with Tool-Augmented Reasoning Agents for Flexible ExtensibilityarXiv2025.6PaperNot Available
The Optimization Paradox in Clinical AI Multi-Agent SystemsarXiv2025.6PaperStar
GitHub
DoctorAgent-RL: A Multi-Agent Collaborative Reinforcement Learning System for Multi-Turn Clinical DialogueEMNLP2025.5PaperStar
GitHub
Silence is Not Consensus: Disrupting Agreement Bias in Multi-Agent LLMs via Catfish Agent for Clinical Decision MakingarXiv2025.5PaperNot Available
MDTeamGPT: A Self-Evolving LLM-Based Multi-Agent Framework for Multi-Disciplinary Team Medical ConsultationEMNLP2025.3PaperStar
GitHub
The Application of MATEC (Multi-AI Agent Team Care) Framework in Sepsis CarearXiv2025.3PaperNot Available
Agentic Medical Knowledge Graphs Enhance Medical Question Answering: Bridging the Gap Between LLMs and Evolving Medical KnowledgeEMNLP Findings2025.2PaperStar
GitHub
A Layered Debating Multi-Agent System for Similar Disease DiagnosisNAACL2025PaperNot Available
KG4Diagnosis: A Hierarchical Multi-Agent LLM Framework with Knowledge Graph EnhancementarXiv2024.12PaperNot Available
Zodiac: A Cardiologist-Level LLM Framework for Multi-Agent DiagnosticsarXiv2024.10PaperNot Available
MedAgents: Large Language Models as Collaborators for Zero-shot Medical ReasoningACL 2024 Findings2023.11PaperStar
GitHub

2. Patient-Facing Applications

2.1 Mental Health & CBT Agents

TitleVenueDatePaper LinkProject Page
ODRA: Synthesizing Cognitive Behavioral Therapy Sessions with Structured Chain-Of-Thought and Dynamic Patient ResistancearXiv2026.08PaperNot Available
Knowledge-augmented Agentic AI for Mental Health Medication Information SeekingarXiv2026.06PaperNot Available
A Multi-Agent Audit Framework for High-Stakes Reasoning: Evaluation and Interpretability in Clinical Mental Health ScreeningarXiv2026.06PaperNot Available
An Agentic LLM-Based Framework for Population-Scale Mental Health ScreeningIEEE BigData2026.05PaperNot Available
AI-Care: A Conversational Agentic System for Task Coordination in Alzheimer's Disease CarearXiv2026.05PaperNot Available
Design and Evaluation of a Culturally Adapted Multimodal Virtual Agent for PTSD ScreeningarXiv2026.04PaperNot Available
OMIND: Framework for Knowledge Grounded Finetuning and Multi-Turn Dialogue Benchmark for Mental Health LLMsarXiv2026.03PaperNot Available
YAQIN: Culturally Sensitive, Agentic AI for Mental Healthcare Support Among Muslim Women in the UKarXiv2026.03PaperNot Available
MIND: Unified Inquiry and Diagnosis RL for Psychiatric ConsultationarXiv2026.03PaperNot Available
SynthAgent: A Multi-Agent LLM Framework for Realistic Patient SimulationAAAI Workshop2026.02PaperNot Available
Advancing AI Trustworthiness Through Patient Simulation for Antidepressant SelectionarXiv2026.02PaperNot Available
DemMA: Dementia Multi-Turn Dialogue Agent with Expert-Guided Reasoning and Action SimulationarXiv2026.01PaperNot Available
CittaVerse (一念万相)AI-powered reminiscence therapy platform for dementia/MCI using narrative identity, autobiographical memory scaffolding, and 6-dimension narrative quality scoringarXiv (in prep)PaperGitHub
coTherapist: A Behavior-Aligned Small Language Model to Support Mental Healthcare ExpertsarXiv2026.01PaperNot Available
Towards Efficient and Robust Linguistic Emotion Diagnosis for Mental HealtharXiv2026.01PaperNot Available
ChatThero: An LLM-Supported Chatbot for Behavior Change and Therapeutic Support in Addiction RecoveryarXiv2025.08PaperStar
GitHub Reproduce
VChatter: Exploring Generative Conversational Agents for Simulating Exposure Therapy to Reduce Social AnxietyarXiv2025.06PaperNot Available
AnnaAgent: Dynamic Evolution Agent System with Multi-Session Memory for Realistic Seeker SimulationACL Findings2025.06PaperStar
GitHub
MIND: Towards Immersive Psychological Healing with Multi-Agent Inner DialogueEMNLP Findings2025.02PaperStar
GitHub Reproduce
Cami: A counselor agent supporting motivational interviewing through state inference and topic explorationACL2025.02PaperStar
GitHub
Autocbt: An autonomous multi-agent framework for cognitive behavioral therapy in psychological counselingarXiv2025.01PaperNot Available
PsyDraw: A Multi-Agent Multimodal System for Mental Health Screening in Left-Behind ChildrenarXiv2024.12PaperStar
GitHub
Cactus: Towards psychological counseling conversations using cognitive behavioral theoryEMNLP Findings2024.07PaperStar
GitHub
MentalAgora: A Gateway to Advanced Personalized Care in Mental Health through Multi-Agent DebatingarXiv2024.07PaperStar
GitHub
Compeer: A generative conversational agent for proactive peer supportarXiv2024.07PaperStar
GitHub
Understanding the benefits and challenges of using large language model-based conversational agents for mental well-being supportAMIA Annual Symposium Proceedings2023.07PaperNot Available

2.2 Clinical Communication & Intake Agents

TitleVenueDatePaper LinkProject Page
Towards Expert-level Medical AI for Real-time Video ConsultationsarXiv2026.08PaperNot Available
Guideline-as-Oracle: Zero-Annotation Training of an Ophthalmic Telephone Triage AgentarXiv2026.08PaperNot Available
Towards Conversational Medical AI with Eyes, Ears and a VoicearXiv2026.05PaperNot Available
SymptomAI: Toward a Conversational AI Agent for Everyday Symptom AssessmentarXiv2026.05PaperNot Available
ClinicBot: A Guideline-Grounded Clinical Chatbot with Prioritized Evidence RAG and Verifiable CitationsarXiv2026.05PaperNot Available
Statistics, Not Scale: Modular Medical Dialogue with Bayesian Belief EnginearXiv2026.04PaperNot Available
EMSDialog: Synthetic Multi-person Emergency Medical Service Dialogue Generation from Electronic Patient Care Reports via Multi-LLM AgentsACL Findings2026.04PaperNot Available
AI Agents for Conversational Patient Triage: Preliminary Simulation-Based Evaluation with Real-World EHR DataarXiv2025.6paperNot Available
A two-stage proactive dialogue generator for efficient clinical information collectionExpert Systems with Applications2025PaperNot Available
PIORS: Personalized Intelligent Outpatient Reception based on Large Language Model with Multi-Agents Medical Scenario SimulationACL Findings2024.11PaperStar
GitHub
A language model--powered simulated patient with automated feedback for history taking: Prospective studyJMIR2024PaperNot Available
Conversational health agents: a personalized large language model-powered agent frameworkJAMIA Open2024PaperStar
GitHub
Talk2Care: Facilitating asynchronous patient-provider communication with large-language-modelarXiv2023.9PaperNot Available

2.3 Screening & Personalized Care Agents

TitleVenueDatePaper LinkProject Page
Traj-Evolve: A Self-Evolving Multi-Agent System for Patient Trajectory Modeling in Lung Cancer Early DetectionarXiv2026.06PaperNot Available
Detecting Clinical Discrepancies in Health Coaching Agents: A Dual-Stream Memory and Reconciliation ArchitecturearXiv2026.04PaperNot Available
Agentic AI for Personalized Physiotherapy: A Multi-Agent Framework for Generative Video Training and Real-Time Pose CorrectionICDH IEEE2026.04PaperNot Available
Sense Less, Infer More: Agentic Multimodal Transformers for Edge Medical IntelligencearXiv2026.04PaperNot Available
NutriOrion: A Hierarchical Multi-Agent Framework for Personalized Nutrition InterventionarXiv2026.02PaperNot Available
FinAgent: An Agentic AI Framework Integrating Personal Finance and Nutrition PlanningarXiv2025.12PaperNot Available
On-device Large Multi-modal Agent for Human Activity RecognitionarXiv2025.12PaperNot Available
Causal Reinforcement Learning based Agent-Patient Interaction with Clinical Domain KnowledgearXiv2025.12PaperNot Available
AI-VaxGuide: An Agentic RAG-Based LLM for Vaccination DecisionsarXiv2025.07Paperhuggingface
A Conversational Agent for Early Detection of Neurotoxic Effects of Medications through Automated Intensive ObservationPACIFIC SYMPOSIUM ON BIOCOMPUTING2024PaperNot Available
An agentic AI tool for personalized body composition screening and metabolic health diagnosticsNot Available2026Not AvailableLink

2.4 General-purpose Healthcare Avatars

TitleVenueDatePaper LinkProject Page
The Anatomy of a Personal Health AgentarXiv2025.08PaperNot Available
A general-purpose AI avatar in healthcarearXiv2024.01PaperNot Available

3. Drug Discovery & Development

TitleVenueDatePaper LinkProject Page
CASCADE: An Agentic Regulatory Network Framework for Patient-Data-Validated Downstream Perturbation PredictionarXiv2026.08PaperNot Available
An AI agent for treatment reasoning over a biomedical tool universearXiv2026.06PaperStar
GitHub
BioInsight: Multi-Agent Orchestration for Interactive Biomedical Knowledge DiscoveryarXiv2026.06PaperNot Available
DeepRoot: A KG-Coordinated Multi-Agent System for Therapeutic Reasoning over Historical Medical TextsarXiv2026.06PaperNot Available
Beyond Prompt-Based Planning: MCP-Native Graph Planning-based Biomedical Agent SystemarXiv2026.06PaperNot Available
A Versatile AI Agent for Rare Disease Diagnosis and Risk Gene Prioritization (Hygieia)arXiv2026.05PaperNot Available
FastOMOP: A Foundational Architecture for Reliable Agentic Real-World Evidence Generation on OMOP CDM dataarXiv2026.04PaperNot Available
Vibe Medicine: Redefining Biomedical Research Through Human-AI Co-WorkarXiv2026.04PaperNot Available
RexDrug: Reliable Multi-Drug Combination Extraction through Reasoning-Enhanced LLMsarXiv2026.03PaperStar
GitHub
ALPACA: A Reinforcement Learning Environment for Medication Repurposing in Alzheimer's DiseasearXiv2026.02PaperNot Available
Causal-Enhanced AI Agents for Medical Research ScreeningarXiv2026.01PaperNot Available
MedAI: Evaluating TxAgent's Therapeutic Agentic Reasoning in the NeurIPS CURE-Bench CompetitionarXiv2025.12PaperBenchmark & Competition
ToolUniverse: An open platform for democratizing AI scientistsarXiv2025.09PaperStar
GitHub
BioScientistAgent: Designing LLM-Biomedical Agents with KG-Augmented RL Reasoning Modulesbiorxiv2025.08PaperNot Available
RAG-Enhanced Collaborative LLM Agents for Drug DiscoveryarXiv2025.02PaperNot Available
Large Language Model Agent for Modular Task Execution in Drug DiscoveryarXiv2025.07PaperStar
GitHub
AUTOCT: Automating Interpretable Clinical Trial Prediction with LLM AgentsarXivEMNLPPaperStar
GitHub
Llm agent swarm for hypothesis-driven drug discoveryarXiv2025.04PaperNot Available
Txgemma: Efficient and agentic llms for therapeuticsarXiv2025.04PaperNot Available
TrialGenie: Empowering Clinical Trial Design with Agentic Intelligence and Real World DatamedRxiv2025.04PaperNot Available
TxAgent: An AI agent for therapeutic reasoning across a universe of toolsarXiv2025.03PaperStar
GitHub
Drugagent: Automating ai-aided drug discovery programming through llm multi-agent collaborationAAAI 2025 workshop AI4Research2024.11PaperStar
GitHub
Drugagent: Multi-agent large language model-based reasoning for drug-target interaction predictionarXiv2024.08PaperStar
GitHub
PRISM: Patient Records Interpretation for Semantic Clinical Trial Matching using Large Language Modelsnpj Digital Medicine2024.01PaperNot Available
MALADE: Orchestration of LLM-powered Agents with Retrieval Augmented Generation for PharmacovigilanceMLHC2024PaperStar
GitHub

4. Healthcare Administration & Workflow

TitleVenueDatePaper LinkProject Page
From Single Chatbots to Governed Agent Ecosystems: An Agentic AI Pattern Catalogue and Orchestration Framework for Mission-Critical Hospital Information Management SystemsarXiv2026.08PaperNot Available
From Siloed Algorithms to Compliance-First Agentic Platforms: A Multi-Layered Architecture for Hospital AI SystemsarXiv2026.08PaperNot Available
Toward Trustworthy Large Language Model Agents in HealthcarearXiv2026.07PaperStar
GitHub
Why Trust Your Agent? Empirical Security Gains from TRiSM-Guided Agentic Workflows in HealthcarearXiv

…(truncated)

Collected info

  • 1,255 stars
  • 159 forks
  • Source updated: 9/22/2026