diff --git a/upstream/AgenticHealthAI-Awesome-AI-Agents-for-Healthcare/catalogue/README.md b/upstream/AgenticHealthAI-Awesome-AI-Agents-for-Healthcare/catalogue/README.md index 12df19f..f432c13 100644 --- a/upstream/AgenticHealthAI-Awesome-AI-Agents-for-Healthcare/catalogue/README.md +++ b/upstream/AgenticHealthAI-Awesome-AI-Agents-for-Healthcare/catalogue/README.md @@ -2,9 +2,9 @@ title: "Awesome AI Agents for Healthcare" task: "" lineage_type: import -upstream_source: https://github.com/AgenticHealthAI/Awesome-AI-Agents-for-Healthcare/blob/02e3270e/README.md -upstream_sha: 02e3270e -imported_at: 2026-07-23 +upstream_source: https://github.com/AgenticHealthAI/Awesome-AI-Agents-for-Healthcare/blob/26bc3511/README.md +upstream_sha: 26bc3511 +imported_at: 2026-08-19 prompt_class: catalogue upstream_changes: accepted author: upstream @@ -99,6 +99,97 @@ If you find our paper and repository helpful, please cite: # Latest Papers ## Year 2026 +1. [arxiv 2026.8] **MARC v1: An Open-Source Multi-Agent Framework for Clinical AI Reasoning and Coordination** [[paper]](https://arxiv.org/abs/2608.13476) [[Github]](https://github.com/Penn-RAIL/MARC-v1) +1. [arxiv 2026.8] **Auditable agentic AI for evidence-grounded thyroid ultrasound diagnosis and reporting** [[paper]](https://arxiv.org/abs/2608.12590) +1. [arxiv 2026.8] **Social Chain of Thought: A Multi-Agent Architecture Grounded in Medical Differential Diagnosis Methodology** [[paper]](https://arxiv.org/abs/2608.11420) +1. [arxiv 2026.8] **Beyond Relevance: Bayesian Evidence Acquisition for Agentic Whole-Slide Image Reasoning** [[paper]](https://arxiv.org/abs/2608.05757) [[Github]](https://github.com/bryanwong17/BEACON) +1. [arxiv 2026.8] **MIRA: Medical Image Reflection for Agentic Diagnosis** [[paper]](https://arxiv.org/abs/2608.10827) +1. [arxiv 2026.8] **Towards Expert-level Medical AI for Real-time Video Consultations** [[paper]](https://arxiv.org/abs/2608.09861) +1. [arxiv 2026.8] **An Agentic Generative Large Language Model for Treatment Planning of Colorectal Cancer** [[paper]](https://arxiv.org/abs/2608.09142) +1. [arxiv 2026.8] **Diagnosing as Cardiologists Do: ECG Agents with Doctor-Grounded Priors for Clinical Reasoning Across Diseases and Populations** [[paper]](https://arxiv.org/abs/2608.09053) +1. [arxiv 2026.8] **ELICITED: EHR-grounded Longitudinal Interactive Conversations for Information-seeking Triage Evaluation and Decision-making** [[paper]](https://arxiv.org/abs/2608.09024) +1. [arxiv 2026.8] **From Single Chatbots to Governed Agent Ecosystems: An Agentic AI Pattern Catalogue and Orchestration Framework for Mission-Critical Hospital Information Management Systems** [[paper]](https://arxiv.org/abs/2608.07627) +1. [arxiv 2026.8] **Does Splitting a Triage Decision Across Agents Hide Bias or Help Catch It? A Multi-Agent Simulation Study of LLM-Based Resource Allocation Under Audit Capacity Constraints** [[paper]](https://arxiv.org/abs/2608.06949) [[Github]](https://github.com/Polpii/policy-town) +1. [arxiv 2026.8] **From Siloed Algorithms to Compliance-First Agentic Platforms: A Multi-Layered Architecture for Hospital AI Systems** [[paper]](https://arxiv.org/abs/2608.06112) +1. [arxiv 2026.8] **DoctorAgents: an agentic framework to iteratively refine AutoML pipeline for small clinical temporal data** [[paper]](https://arxiv.org/abs/2608.05375) +1. [arxiv 2026.8] **CASCADE: An Agentic Regulatory Network Framework for Patient-Data-Validated Downstream Perturbation Prediction** [[paper]](https://arxiv.org/abs/2608.05359) +1. [arxiv 2026.8] **Guideline-as-Oracle: Zero-Annotation Training of an Ophthalmic Telephone Triage Agent** [[paper]](https://arxiv.org/abs/2608.04772) +1. [arxiv 2026.8] **ODRA: Synthesizing Cognitive Behavioral Therapy Sessions with Structured Chain-Of-Thought and Dynamic Patient Resistance** [[paper]](https://arxiv.org/abs/2608.04524) +1. [arxiv 2026.8] **Agents Catching Agents: Shortcut Cascades and Benchmark Gaming in Clinical Multi-Agent Systems** [[paper]](https://arxiv.org/abs/2608.03744) [[Github]](https://github.com/criticaldata/benchmaxxing) +1. [arxiv 2026.8] **TumorBoard: Evidence-Grounded Multi-Agent Decision Support for Longitudinal Neuro-Oncology** [[paper]](https://arxiv.org/abs/2608.03190) +1. [arxiv 2026.7] **CyberNeuro: A Privacy-Preserving Agentic Workbench for Cohort-Scale Neuroimage and Clinical Data Analysis** [[paper]](https://arxiv.org/abs/2607.28841) +1. [CVPR 2026 Workshop] **Inference-Time Agentic Decision Rules Beat Longer Evolving Search for Multi-Image Medical Reasoning** [[paper]](https://arxiv.org/abs/2607.27564) +1. [arxiv 2026.7] **ClinLens: Towards Long-Horizon Coding Agents for Longitudinal Multimodal Clinical Data Science** [[paper]](https://arxiv.org/abs/2607.26155) +1. [arxiv 2026.7] **Agentic AI in medicine: architectures, applications, evaluation, and challenges for clinical translation** [[paper]](https://arxiv.org/abs/2607.25489) +1. [arxiv 2026.7] **PatientAgentBench: A Benchmark Framework for Evaluating Patient-Facing Health AI Agents** [[paper]](https://arxiv.org/abs/2607.25485) [[Github]](https://github.com/amazon-science/PatientAgentBench) +1. [arxiv 2026.7] **Cardiologent: Multi-Agent Clinical Decision Support for Patient-Level Arrhythmia Assessment, Urgency, and Management** [[paper]](https://arxiv.org/abs/2607.25340) +1. [arxiv 2026.7] **Spectral Dynamics of Semantic Drift in Clinical Multi-Agent Language Model Networks** [[paper]](https://arxiv.org/abs/2607.22758) +1. [MICCAI 2026] **Agent-Guided Relational Concept Discovery: Toward Interpretable Surgical Margin Assessment** [[paper]](https://arxiv.org/abs/2607.21437) +1. [arxiv 2026.7] **Bayesian uncertainty estimation improves clinical decision making in medical AI agents** [[paper]](https://arxiv.org/abs/2607.20582) +1. [arxiv 2026.7] **Harnessing Disagreement: Detecting Correlated Agreement Blindness in Multi-Agent Triage** [[paper]](https://arxiv.org/abs/2607.19899) +1. [arxiv 2026.7] **MedDDC-Eval: Diagnosis-Decoupled Evaluation of Multi-Turn Medical Consultation Agents** [[paper]](https://arxiv.org/abs/2607.18999) +1. [arxiv 2026.7] **Understanding From Human Perspective: A Multi-agent System for Interactive Egocentric Medical Image Segmentation** [[paper]](https://arxiv.org/abs/2607.17341) [[Github]](https://github.com/wdyyyyyy/EgoMed-Agent) +1. [arxiv 2026.7] **Cura 1T: Specialized Model for Agentic Healthcare** [[paper]](https://arxiv.org/abs/2607.15314) [[Github]](https://github.com/actava-ai/Cura) +1. [arxiv 2026.7] **Evaluating Frontier AI Agents as Autonomous Clinical Security Auditors** [[paper]](https://arxiv.org/abs/2607.13411) +1. [arxiv 2026.7] **A Multi-Agent System for Autonomous, Fine-Tuning-Free Clinical Symptom Detection: Development and Validation Study** [[paper]](https://arxiv.org/abs/2607.12886) +1. [arxiv 2026.7] **Agentic systems for breast cancer treatment recommendations** [[paper]](https://arxiv.org/abs/2607.12051) [[Github]](https://github.com/GRUPOMED4U/breast_cancer_agents_paper) +1. [arxiv 2026.7] **The Path to Self-Evolving Clinical Systems: Scaling Medical Agents from Assistance to Autonomy** [[paper]](https://arxiv.org/abs/2607.11175) [[Github]](https://github.com/zhcz328/Awesome-Medical-Agents) +1. [arxiv 2026.7] **Policy-Driven CT-Agent: Modeling Phase-Aware Diagnostic Control for Clinically Consistent CT Reasoning** [[paper]](https://arxiv.org/abs/2607.10748) +1. [arxiv 2026.7] **Towards Autonomous and Auditable Medical Imaging Model Development** [[paper]](https://arxiv.org/abs/2607.10522) +1. [arxiv 2026.7] **Information-seeking failures of large language models in agentic clinical reasoning** [[paper]](https://arxiv.org/abs/2607.10275) +1. [arxiv 2026.7] **LongMedBench: Benchmarking Medical Agents for Long-Horizon Clinical Decision-Making** [[paper]](https://arxiv.org/abs/2607.09322) +1. [arxiv 2026.7] **Trust but Verify:Evidence-Linked Multi-Agent Clinical Information Extraction in Pathology** [[paper]](https://arxiv.org/abs/2607.06435) +1. [arxiv 2026.7] **Toward Trustworthy Large Language Model Agents in Healthcare** [[paper]](https://arxiv.org/abs/2607.05055) [[Github]](https://github.com/Hadi-Hsn/CareConnect) +1. [arxiv 2026.7] **Multi-Large Language Model Orchestrated Severity Assessment of Clinical Records (MOSAIC)** [[paper]](https://arxiv.org/abs/2607.05032) +1. [arxiv 2026.7] **CogRad: A Cognitively-Inspired Multi-Agent Framework for Radiology Report Generation** [[paper]](https://arxiv.org/abs/2607.03853) +1. [arxiv 2026.7] **MedCalc-Pro: Solving Complex Medical Calculations with LLM Agents** [[paper]](https://arxiv.org/abs/2607.02879) +1. [arxiv 2026.7] **Evaluating Agentic Harness Systems for Autonomous Computational Pathology** [[paper]](https://arxiv.org/abs/2607.02598) +1. [arxiv 2026.6] **HealthAgentBench: A Unified Benchmark Suite of Realistic Agentic Healthcare Environments for Challenging Frontier AI Agents** [[paper]](https://arxiv.org/abs/2606.31179) [[Github]](https://github.com/microsoft/HealthAgentBench) +1. [AMIA 2026] **Agentic AI Enhances Physician Trust in Clinical Decision Making** [[paper]](https://arxiv.org/abs/2606.30658) +1. [arxiv 2026.6] **TopoAgent: An Agentic Framework for Automated Topology Learning in Medical Imaging** [[paper]](https://arxiv.org/abs/2606.29763) +1. [IJCAI 2026] **DEEPMED Search: An Open-Source Agentic Platform for Medical Deep Research with Introspective Verification** [[paper]](https://arxiv.org/abs/2606.29746) +1. [arxiv 2026.6] **MedEvoEval: Evaluating Continual Evolution of Doctor Agents through Simulated Clinical Episodes** [[paper]](https://arxiv.org/abs/2606.28900) +1. [arxiv 2026.6] **An AI agent for treatment reasoning over a biomedical tool universe** [[paper]](https://arxiv.org/abs/2606.28692) [[Github]](https://github.com/mims-harvard/ATHENA) +1. [arxiv 2026.6] **Why Trust Your Agent? Empirical Security Gains from TRiSM-Guided Agentic Workflows in Healthcare** [[paper]](https://arxiv.org/abs/2606.28666) +1. [MICCAI 2026] **CPAgents: Agentic Composite Phenotype Generation for Cardiac Disease Association** [[paper]](https://arxiv.org/abs/2606.28179) +1. [arxiv 2026.6] **Knowledge-augmented Agentic AI for Mental Health Medication Information Seeking** [[paper]](https://arxiv.org/abs/2606.26205) +1. [arxiv 2026.6] **MedGuards: Multi-Agent System for Reliable Medical Error Detection and Correction** [[paper]](https://arxiv.org/abs/2606.25651) [[Github]](https://github.com/congboma/MedGuards) +1. [arxiv 2026.6] **DeepBD: A Grounded Agentic Workflow for Variant Prioritization and Diagnosis of Genetic Birth Defects** [[paper]](https://arxiv.org/abs/2606.24779) +1. [arxiv 2026.6] **EHR-Complex: Benchmarking Medical Agents for Complex Clinical Reasoning** [[paper]](https://arxiv.org/abs/2606.23301) +1. [MICCAI 2026] **Evo-RAD: Navigating Rare Retinal Disease Diagnosis via Self-Evolving Agentic Retrieval** [[paper]](https://arxiv.org/abs/2606.22955) [[Github]](https://github.com/SDH-Lab/Evo-RAD) +1. [arxiv 2026.6] **OpenBioRQ: Unsolved Biomedical Research Questions for Agents** [[paper]](https://arxiv.org/abs/2606.21959) +1. [arxiv 2026.6] **A Multi-Agent Audit Framework for High-Stakes Reasoning: Evaluation and Interpretability in Clinical Mental Health Screening** [[paper]](https://arxiv.org/abs/2606.21123) +1. [arxiv 2026.6] **BioInsight: Multi-Agent Orchestration for Interactive Biomedical Knowledge Discovery** [[paper]](https://arxiv.org/abs/2606.20997) +1. [arxiv 2026.6] **Democratizing and accelerating AI-driven pathology research through agentic intelligence** [[paper]](https://arxiv.org/abs/2606.20677) +1. [arxiv 2026.6] **MedRLM: Recursive Multimodal Health Intelligence for Long-Context Clinical Reasoning, Sensor-Guided Screening, Evidence-Grounded Decision Support, and Community-to-Tertiary Referral Optimization** [[paper]](https://arxiv.org/abs/2606.20164) +1. [arxiv 2026.6] **Prompt, Plan, Extract: Zero-Shot Agentic LLMs Workflows for Lung Pathology Extraction from Clinical Narratives** [[paper]](https://arxiv.org/abs/2606.19852) +1. [arxiv 2026.6] **Configurable Clinical Information Extraction with Agentic RAG: What Works, What Breaks, and Why** [[paper]](https://arxiv.org/abs/2606.19602) +1. [arxiv 2026.6] **Are LLMs Ready to Assist Physicians? PhysAssistBench for Interactive Doctor-Patient-EHR Assistance** [[paper]](https://arxiv.org/abs/2606.18613) +1. [arxiv 2026.6] **RubricsTree: Scalable and Evolving Open-Ended Evaluation of Personal Health Agents across Health Memory and Medical Skills** [[paper]](https://arxiv.org/abs/2606.18203) +1. [arxiv 2026.6] **Agentic AI-based Framework for Mitigating Premature Diagnostic Handoff and Silent Hallucination in Healthcare Applications** [[paper]](https://arxiv.org/abs/2606.18068) +1. [arxiv 2026.6] **MedEasy: Designing AI Standardized Patients for Clinical Consultation Training** [[paper]](https://arxiv.org/abs/2606.17512) +1. [arxiv 2026.6] **Teaching agentic AI to learn expert reasoning for rare disease diagnosis** [[paper]](https://arxiv.org/abs/2606.16149) +1. [arxiv 2026.6] **DeepRoot: A KG-Coordinated Multi-Agent System for Therapeutic Reasoning over Historical Medical Texts** [[paper]](https://arxiv.org/abs/2606.15931) +1. [arxiv 2026.6] **Let LLMs Judge Each Other: Multi-Agent Peer-Reviewed Reasoning for Medical Question Answering** [[paper]](https://arxiv.org/abs/2606.15419) +1. [arxiv 2026.6] **XMedFusion: A Knowledge-Guided Multimodal Perception and Reasoning Framework for Autonomous Medical Systems** [[paper]](https://arxiv.org/abs/2606.14766) +1. [arxiv 2026.6] **Trust but Verify: Mitigating Medical Hallucinations via Post-Hoc Adversarial Auditing and Multi-Agent Feedback Loops** [[paper]](https://arxiv.org/abs/2606.14149) +1. [arxiv 2026.6] **MedLatentDx: Latent Multi-Agent Communication for Cross-Hospital Rare-Disease Diagnosis** [[paper]](https://arxiv.org/abs/2606.13945) +1. [IJCAI 2026] **ArogyaSutra: A Multi-Agent Framework for Multimodal Medical Reasoning in Indic Languages** [[paper]](https://arxiv.org/abs/2606.13572) +1. [arxiv 2026.6] **Skill-Augmented AI Agents for Medical Research Analysis: An Exploratory Multi-Model Human Evaluation in an NSCLC Transcriptomic Biomarker Task** [[paper]](https://arxiv.org/abs/2606.11830) +1. [arxiv 2026.6] **MedCTA: A Benchmark for Clinical Tool Agents** [[paper]](https://arxiv.org/abs/2606.11702) [[Github]](https://github.com/IVUL-KAUST/MedCTA) +1. [arxiv 2026.6] **Experience Makes Skillful: Enabling Generalizable Medical Agent Reasoning via Self-Evolving Skill Memory** [[paper]](https://arxiv.org/abs/2606.09365) +1. [arxiv 2026.6] **Baichuan-M4: A Clinical-Grade Medical Agent System for Continuous Care** [[paper]](https://arxiv.org/abs/2606.08982) +1. [arxiv 2026.6] **A multi-agent system for spine MRI report generation from multi-sequence imaging** [[paper]](https://arxiv.org/abs/2606.08897) +1. [arxiv 2026.6] **A Multi-modal Agentic Co-pilot for Evidence Grounded Computational Pathology** [[paper]](https://arxiv.org/abs/2606.08093) +1. [arxiv 2026.6] **PSEBench: A Controllable and Verifiable Benchmark for Evaluating LLMs in Patient Safety Event Triage** [[paper]](https://arxiv.org/abs/2606.05463) +1. [arxiv 2026.6] **Beyond Prompt-Based Planning: MCP-Native Graph Planning-based Biomedical Agent System** [[paper]](https://arxiv.org/abs/2606.04494) +1. [arxiv 2026.6] **D2MDT: Department-aware Multidisciplinary Team Consultation with Deliberation for Efficient Clinical Prediction** [[paper]](https://arxiv.org/abs/2606.03543) +1. [arxiv 2026.6] **MeDxAgent: Multi-Agent Consultation for Interactive Medical Diagnosis** [[paper]](https://arxiv.org/abs/2606.03416) +1. [arxiv 2026.6] **MedCUA-Bench: A Screenshot-Only Benchmark for Clinical Computer-Use Agents** [[paper]](https://arxiv.org/abs/2606.03203) +1. [arxiv 2026.6] **ClinicalMC: A Benchmark for Multi-Course Clinical Decision-Making with Large Language Models** [[paper]](https://arxiv.org/abs/2606.03157) +1. [arxiv 2026.6] **Traj-Evolve: A Self-Evolving Multi-Agent System for Patient Trajectory Modeling in Lung Cancer Early Detection** [[paper]](https://arxiv.org/abs/2606.02812) +1. [arxiv 2026.6] **ClinEnv: An Interactive Multi-Stage Long Horizon EHR Environment for Agents** [[paper]](https://arxiv.org/abs/2606.02568) +1. [arxiv 2026.6] **AutoMedBench: Towards Medical AutoResearch with Agentic AI Models** [[paper]](https://arxiv.org/abs/2606.01961) 1. [arxiv 2026.5] **SEMA-RAG: A Self-Evolving Multi-Agent Retrieval-Augmented Generation Framework for Medical Reasoning** [[paper]](https://arxiv.org/abs/2605.17101) 1. [arxiv 2026.5] **CHI-Bench: Can AI Agents Automate End-to-End, Long-Horizon, Policy-Rich Healthcare Workflows?** [[paper]](https://arxiv.org/abs/2605.16679) [[Github]](https://github.com/actava-ai/chi-bench) [[Project]](https://actava.ai/benchmarks) 1. [arxiv 2026.5] **COTCAgent: Preventive Consultation via Probabilistic Chain-of-Thought Completion** [[paper]](https://arxiv.org/abs/2605.15016) @@ -357,7 +448,6 @@ If you find our paper and repository helpful, please cite: 1. [arxiv 2025.8] **Patho-AgenticRAG: Towards Multimodal Agentic Retrieval-Augmented Generation for Pathology VLMs via Reinforcement Learning** [[paper]](http://arxiv.org/abs/2508.02258v1) [[code]](https://github.com/Wenchuan-Zhang/Patho-AgenticRAG) 1. [arxiv 2025.8] **Agent-Based Feature Generation from Clinical Notes for Outcome Prediction** [[paper]](http://arxiv.org/abs/2508.01956v1) 1. [arxiv 2025.8] **GMAT: Grounded Multi-Agent Clinical Description Generation for Text Encoder in Vision-Language MIL for Whole Slide Image Classification** [[paper]](http://arxiv.org/abs/2508.01293v1) -1. [arxiv 2025.8] **A Multi-Agent Approach to Neurological Clinical Reasoning** [[paper]](https://arxiv.org/abs/2508.14063) 1. [biorxiv 2025.8] **BioScientistAgent: Designing LLM-Biomedical Agents with KG-Augmented RL Reasoning Modules for Drug Repurposing and Mechanistic of Action Elucidation** [[paper]](https://www.biorxiv.org/content/10.1101/2025.08.08.669291) 1. [arxiv 2025.7] **Agentic AI framework for end-to-end medical data inference** [[paper]](http://arxiv.org/abs/2507.18115v1) 1. [arxiv 2025.7] **Resilient Multi-Agent Negotiation for Medical Supply Chains: Integrating LLMs and Blockchain for Transparent Coordination** [[paper]](http://arxiv.org/abs/2507.17134v1) @@ -597,6 +687,12 @@ _(Agents designed to process and reason over multiple data types like images, te | Title | Venue | Date | Paper Link | Project Page | | :---------------------------------------------------------------------------------------------------------- | :------ | :------ | :---------------------------------------------------------------------------------------------------------------------------- | :--------------------------------------------------------------------------------------------------------------------------------------------------------- | +| **MIRA: Medical Image Reflection for Agentic Diagnosis** | arXiv | 2026.08 | [Paper](https://arxiv.org/abs/2608.10827) | Not Available | +| **Inference-Time Agentic Decision Rules Beat Longer Evolving Search for Multi-Image Medical Reasoning** | CVPR Workshop | 2026.07 | [Paper](https://arxiv.org/abs/2607.27564) | Not Available | +| **Understanding From Human Perspective: A Multi-agent System for Interactive Egocentric Medical Image Segmentation** | arXiv | 2026.07 | [Paper](https://arxiv.org/abs/2607.17341) | ![Star](https://img.shields.io/github/stars/wdyyyyyy/EgoMed-Agent.svg?style=social&label=Star)
[GitHub](https://github.com/wdyyyyyy/EgoMed-Agent) | +| **MedRLM: Recursive Multimodal Health Intelligence for Long-Context Clinical Reasoning, Sensor-Guided Screening, Evidence-Grounded Decision Support, and Community-to-Tertiary Referral Optimization** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.20164) | Not Available | +| **XMedFusion: A Knowledge-Guided Multimodal Perception and Reasoning Framework for Autonomous Medical Systems** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.14766) | Not Available | +| **ArogyaSutra: A Multi-Agent Framework for Multimodal Medical Reasoning in Indic Languages** | IJCAI | 2026.06 | [Paper](https://arxiv.org/abs/2606.13572) | Not Available | | **Towards Conversational Medical AI with Eyes, Ears and a Voice** | arXiv | 2026.05 | [Paper](https://arxiv.org/abs/2605.09272) | Not Available | | **VERITAS: A Multi-Agent Co-Scientist for Verifiable Image-Derived Hypothesis Testing** | arXiv | 2026.04 | [Paper](https://arxiv.org/abs/2604.12144) | ![Star](https://img.shields.io/github/stars/LucZot/veritas.svg?style=social&label=Star)
[GitHub](https://github.com/LucZot/veritas) | | **Camyla: Scaling Autonomous Research in Medical Image Segmentation** | arXiv | 2026.04 | [Paper](https://arxiv.org/abs/2604.10696) | [Project](https://yifangao112.github.io/camyla-page/) | @@ -629,6 +725,9 @@ _(Agents designed to process and reason over multiple data types like images, te | Title | Venue | Date | Paper Link | Project Page | | :----------------------------------------------------------------------------------------------------------------------- | :----------------------------------------- | :------ | :-------------------------------------------------------------------- | :----------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- | +| **Policy-Driven CT-Agent: Modeling Phase-Aware Diagnostic Control for Clinically Consistent CT Reasoning** | arXiv | 2026.07 | [Paper](https://arxiv.org/abs/2607.10748) | Not Available | +| **CogRad: A Cognitively-Inspired Multi-Agent Framework for Radiology Report Generation** | arXiv | 2026.07 | [Paper](https://arxiv.org/abs/2607.03853) | Not Available | +| **A multi-agent system for spine MRI report generation from multi-sequence imaging** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.08897) | Not Available | | **ABRA: Agent Benchmark for Radiology Applications** | arXiv | 2026.05 | [Paper](https://arxiv.org/abs/2605.11224) | Not Available | | **DeepTumorVQA: A Hierarchical 3D CT Benchmark for Stage-Wise Evaluation of Medical VLMs and Tool-Augmented Agents** | arXiv | 2026.05 | [Paper](https://arxiv.org/abs/2605.09679) | Not Available | | **GAZE: Grounded Agentic Zero-shot Evaluation with Viewer-Level Tools and Literature Retrieval on Rare Brain MRI** | arXiv | 2026.05 | [Paper](https://arxiv.org/abs/2605.00876) | Not Available | @@ -664,6 +763,10 @@ _(Agents designed to process and reason over multiple data types like images, te | Title | Venue | Date | Paper Link | Project Page | | :---------------------------------------------------------------------------------------------------------------- | :------------ | :------ | :-------------------------------------------------------- | :------------------------------------------------------------------------------------------------------------------------------------------------------------------------- | +| **Trust but Verify:Evidence-Linked Multi-Agent Clinical Information Extraction in Pathology** | arXiv | 2026.07 | [Paper](https://arxiv.org/abs/2607.06435) | Not Available | +| **Democratizing and accelerating AI-driven pathology research through agentic intelligence** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.20677) | Not Available | +| **Prompt, Plan, Extract: Zero-Shot Agentic LLMs Workflows for Lung Pathology Extraction from Clinical Narratives** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.19852) | Not Available | +| **A Multi-modal Agentic Co-pilot for Evidence Grounded Computational Pathology** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.08093) | Not Available | | **Computational Pathology in the Era of Emerging Foundation and Agentic AI -- International Expert Perspectives** | arXiv | 2026.03 | [Paper](https://arxiv.org/abs/2603.05884) | Not Available | | **LAMMI-Pathology: A Tool-Centric Bottom-Up LVLM-Agent Framework for Molecularly Informed Medical Intelligence** | arXiv | 2026.02 | [Paper](https://arxiv.org/abs/2602.18773) | Not Available | | **SurvAgent: Hierarchical CoT-Enhanced Case Banking and Dichotomy-Based Multi-Agent System for Multimodal Survival Prediction** | arXiv | 2025.11 | [Paper](https://arxiv.org/abs/2511.16635) | Not Available | @@ -682,6 +785,8 @@ _(Agents designed to process and reason over multiple data types like images, te | Title | Venue | Date | Paper Link | Project Page | | :---------------------------------------------------------------------- | :----- | :------ | :----------------------------------------- | :------------------------------------------------------------------------------------------------------------------------------------------------- | +| **Diagnosing as Cardiologists Do: ECG Agents with Doctor-Grounded Priors for Clinical Reasoning Across Diseases and Populations** | arXiv | 2026.08 | [Paper](https://arxiv.org/abs/2608.09053) | Not Available | +| **Cardiologent: Multi-Agent Clinical Decision Support for Patient-Level Arrhythmia Assessment, Urgency, and Management** | arXiv | 2026.07 | [Paper](https://arxiv.org/abs/2607.25340) | Not Available | | **ECG Foundation Models and Medical LLMs for Agentic Cardiovascular Intelligence at the Edge: A Review and Outlook** | arXiv | 2026.04 | [Paper](https://arxiv.org/abs/2604.02501) | Not Available | | **Multi-Agent Reasoning for Cardiovascular Imaging Phenotype Analysis** | MICCAI | 2025.07 | [Paper](http://arxiv.org/abs/2507.03460v2) | ![Star](https://img.shields.io/github/stars/MengyunQ/MESHAgents.svg?style=social&label=Star)
[GitHub](https://github.com/MengyunQ/MESHAgents) | @@ -689,6 +794,7 @@ _(Agents designed to process and reason over multiple data types like images, te | Title | Venue | Date | Paper Link | Project Page | | :----------------------------------------------------------------------------------------- | :---- | :------ | :----------------------------------------- | :------------------------------------------------------------------------------------------------------------------------------- | +| **Auditable agentic AI for evidence-grounded thyroid ultrasound diagnosis and reporting** | arXiv | 2026.08 | [Paper](https://arxiv.org/abs/2608.12590) | Not Available | | **Echo-α: Large Agentic Multimodal Reasoning Model for Ultrasound Interpretation** | arXiv | 2026.04 | [Paper](https://arxiv.org/abs/2604.28011) | Not Available | | **Anatomical Prior-Driven Framework for Autonomous Robotic Cardiac Ultrasound Standard View Acquisition** | ICRA | 2026.03 | [Paper](https://arxiv.org/abs/2603.21134) | Not Available | | **Intelligent Virtual Sonographer (IVS): Enhancing Physician-Robot-Patient Communication** | arXiv | 2025.07 | [Paper](http://arxiv.org/abs/2507.13052v1) | ![Star](https://img.shields.io/github/stars/stytim/IVS.svg?style=social&label=Star)
[GitHub](https://github.com/stytim/IVS) | @@ -718,6 +824,8 @@ _(Agents designed to process and reason over multiple data types like images, te | Title | Venue | Date | Paper Link | Project Page | | :----------------------------------------------------------------------------------------------------- | :------------- | :------ | :---------------------------------------------------------------------------- | :--------------------------------------------------------------------------------------------------------------------------------------------------------------------------- | +| **CPAgents: Agentic Composite Phenotype Generation for Cardiac Disease Association** | MICCAI | 2026.06 | [Paper](https://arxiv.org/abs/2606.28179) | Not Available | +| **DeepBD: A Grounded Agentic Workflow for Variant Prioritization and Diagnosis of Genetic Birth Defects** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.24779) | Not Available | | **Autonomous Agent-Orchestrated Digital Twins (AADT): State Synchronization in Rare Genetic Disorders** | arXiv | 2026.03 | [Paper](https://arxiv.org/abs/2603.27104) | Not Available | | **ProtRLSearch: A Multi-Round Multimodal Protein Search Agent with LLMs Trained via RL** | arXiv | 2026.03 | [Paper](https://arxiv.org/abs/2603.01464) | Not Available | | **Geneagent: self-verification language agent for gene-set analysis using domain databases** | Nature Methods | 2025 | [Paper](https://doi.org/10.1038/s41592-025-02748-6) | ![Star](https://img.shields.io/github/stars/ncbi-nlp/GeneAgent.svg?style=social&label=Star)
[GitHub](https://github.com/ncbi-nlp/GeneAgent) | @@ -730,6 +838,9 @@ _(Agents designed to process and reason over multiple data types like images, te | Title | Venue | Date | Paper Link | Project Page | | :--------------------------------------------------------------------------------------------------------------------------- | :------------------- | :------ | :------------------------------------------------------------------------------------------------------ | :--------------------------------------------------------------------------------------------------------------------------------------------------------------- | +| **A Multi-Agent System for Autonomous, Fine-Tuning-Free Clinical Symptom Detection: Development and Validation Study** | arXiv | 2026.07 | [Paper](https://arxiv.org/abs/2607.12886) | Not Available | +| **Multi-Large Language Model Orchestrated Severity Assessment of Clinical Records (MOSAIC)** | arXiv | 2026.07 | [Paper](https://arxiv.org/abs/2607.05032) | Not Available | +| **Configurable Clinical Information Extraction with Agentic RAG: What Works, What Breaks, and Why** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.19602) | Not Available | | **COTCAgent: Preventive Consultation via Probabilistic Chain-of-Thought Completion** | arXiv | 2026.05 | [Paper](https://arxiv.org/abs/2605.15016) | Not Available | | **Reinforcement Learning for Tool-Calling Agents in Fast Healthcare Interoperability Resources (FHIR)** | arXiv | 2026.05 | [Paper](https://arxiv.org/abs/2605.14126) | Not Available | | **Generating synthetic electronic health record data using agent-based models to evaluate machine learning robustness under mass casualty incidents** | CHIL | 2026.05 | [Paper](https://arxiv.org/abs/2605.09951) | Not Available | @@ -767,6 +878,7 @@ _(Agents designed to process and reason over multiple data types like images, te | Title | Venue | Date | Paper Link | Project Page | | :--------------------------------------------------------------------------------------------------------------------- | :----- | :----- | :---------------------------------------------------------------------- | :----------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- | +| **Agent-Guided Relational Concept Discovery: Toward Interpretable Surgical Margin Assessment** | MICCAI | 2026.07 | [Paper](https://arxiv.org/abs/2607.21437) | Not Available | | **CSAP-Assist: Instrument-Agent Dialogue Empowered Vision-Language Models for Collaborative Surgical Action Planning** | MICCAI | 2025 | [Paper](https://link.springer.com/chapter/10.1007/978-3-032-05114-1_14) | ![Star](https://img.shields.io/github/stars/einnullnull/Collaborative-Surgical-Action-Planning-Assist.svg?style=social&label=Star)
[GitHub](https://github.com/einnullnull/Collaborative-Surgical-Action-Planning-Assist) | | **Privacy-Preserving Operating Room Workflow Analysis using Digital Twins** | arXiv | 2025.4 | [Paper](https://arxiv.org/abs/2504.12552) | Not Available | @@ -774,6 +886,7 @@ _(Agents designed to process and reason over multiple data types like images, te | Title | Venue | Date | Paper Link | Project Page | | :--------------------------------------------------------------------------------------------------------------------- | :----- | :----- | :---------------------------------------------------------------------- | :----------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- | +| **MedEasy: Designing AI Standardized Patients for Clinical Consultation Training** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.17512) | Not Available | | **Rethinking Patient Education as Multi-turn Multi-modal Interaction** | arXiv | 2026.04 | [Paper](https://arxiv.org/abs/2604.14656) | Not Available | | **Persona-Based Requirements Engineering for Explainable Multi-Agent Educational Systems: A Scenario Simulator for Clinical Reasoning Training** | CSTE | 2026.04 | [Paper](https://arxiv.org/abs/2604.17186) | Not Available | | **Dialogue to Question Generation for Evidence-based Medical Guideline Agent Development** | ML4H | 2026.03 | [Paper](https://arxiv.org/abs/2603.23937) | Not Available | @@ -785,6 +898,22 @@ _(Agents designed to process and reason over multiple data types like images, te | Title | Venue | Date | Paper Link | Project Page | | :------------------------------------------------------------------------------------------------------------------------------------ | :---------------- | :------ | :----------------------------------------------------- | :--------------------------------------------------------------------------------------------------------------------------------------------------------------- | +| **MARC v1: An Open-Source Multi-Agent Framework for Clinical AI Reasoning and Coordination** | arXiv | 2026.08 | [Paper](https://arxiv.org/abs/2608.13476) | ![Star](https://img.shields.io/github/stars/Penn-RAIL/MARC-v1.svg?style=social&label=Star)
[GitHub](https://github.com/Penn-RAIL/MARC-v1) | +| **Social Chain of Thought: A Multi-Agent Architecture Grounded in Medical Differential Diagnosis Methodology** | arXiv | 2026.08 | [Paper](https://arxiv.org/abs/2608.11420) | Not Available | +| **Harnessing Disagreement: Detecting Correlated Agreement Blindness in Multi-Agent Triage** | arXiv | 2026.07 | [Paper](https://arxiv.org/abs/2607.19899) | Not Available | +| **MedCalc-Pro: Solving Complex Medical Calculations with LLM Agents** | arXiv | 2026.07 | [Paper](https://arxiv.org/abs/2607.02879) | Not Available | +| **DEEPMED Search: An Open-Source Agentic Platform for Medical Deep Research with Introspective Verification** | IJCAI | 2026.06 | [Paper](https://arxiv.org/abs/2606.29746) | Not Available | +| **MedGuards: Multi-Agent System for Reliable Medical Error Detection and Correction** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.25651) | ![Star](https://img.shields.io/github/stars/congboma/MedGuards.svg?style=social&label=Star)
[GitHub](https://github.com/congboma/MedGuards) | +| **Evo-RAD: Navigating Rare Retinal Disease Diagnosis via Self-Evolving Agentic Retrieval** | MICCAI | 2026.06 | [Paper](https://arxiv.org/abs/2606.22955) | ![Star](https://img.shields.io/github/stars/SDH-Lab/Evo-RAD.svg?style=social&label=Star)
[GitHub](https://github.com/SDH-Lab/Evo-RAD) | +| **Agentic AI-based Framework for Mitigating Premature Diagnostic Handoff and Silent Hallucination in Healthcare Applications** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.18068) | Not Available | +| **Teaching agentic AI to learn expert reasoning for rare disease diagnosis** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.16149) | Not Available | +| **Let LLMs Judge Each Other: Multi-Agent Peer-Reviewed Reasoning for Medical Question Answering** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.15419) | Not Available | +| **Trust but Verify: Mitigating Medical Hallucinations via Post-Hoc Adversarial Auditing and Multi-Agent Feedback Loops** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.14149) | Not Available | +| **MedLatentDx: Latent Multi-Agent Communication for Cross-Hospital Rare-Disease Diagnosis** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.13945) | Not Available | +| **Experience Makes Skillful: Enabling Generalizable Medical Agent Reasoning via Self-Evolving Skill Memory** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.09365) | Not Available | +| **Baichuan-M4: A Clinical-Grade Medical Agent System for Continuous Care** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.08982) | Not Available | +| **D2MDT: Department-aware Multidisciplinary Team Consultation with Deliberation for Efficient Clinical Prediction** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.03543) | Not Available | +| **MeDxAgent: Multi-Agent Consultation for Interactive Medical Diagnosis** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.03416) | Not Available | | **SEMA-RAG: A Self-Evolving Multi-Agent Retrieval-Augmented Generation Framework for Medical Reasoning** | arXiv | 2026.05 | [Paper](https://arxiv.org/abs/2605.17101) | Not Available | | **MedExAgent: Training LLM Agents to Ask, Examine, and Diagnose in Noisy Clinical Environments** | arXiv | 2026.05 | [Paper](https://arxiv.org/abs/2605.07058) | Not Available | | **Thinking Like a Clinician: A Cognitive AI Agent for Clinical Diagnosis via Panoramic Profiling and Adversarial Debate** | arXiv | 2026.04 | [Paper](https://arxiv.org/abs/2604.23605) | Not Available | @@ -863,6 +992,9 @@ _(Agents designed to process and reason over multiple data types like images, te | Title | Venue | Date | Paper Link | Project Page | | :------------------------------------------------------------------------------------------------------------------------------------ | :-------------------------------- | :------ | :----------------------------------------- | :--------------------------------------------------------------------------------------------------------------------------------------------------------------------------- | +| **ODRA: Synthesizing Cognitive Behavioral Therapy Sessions with Structured Chain-Of-Thought and Dynamic Patient Resistance** | arXiv | 2026.08 | [Paper](https://arxiv.org/abs/2608.04524) | Not Available | +| **Knowledge-augmented Agentic AI for Mental Health Medication Information Seeking** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.26205) | Not Available | +| **A Multi-Agent Audit Framework for High-Stakes Reasoning: Evaluation and Interpretability in Clinical Mental Health Screening** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.21123) | Not Available | | **An Agentic LLM-Based Framework for Population-Scale Mental Health Screening** | IEEE BigData | 2026.05 | [Paper](https://arxiv.org/abs/2605.13046) | Not Available | | **AI-Care: A Conversational Agentic System for Task Coordination in Alzheimer's Disease Care** | arXiv | 2026.05 | [Paper](https://arxiv.org/abs/2605.08480) | Not Available | | **Design and Evaluation of a Culturally Adapted Multimodal Virtual Agent for PTSD Screening** | arXiv | 2026.04 | [Paper](https://arxiv.org/abs/2604.17871) | Not Available | @@ -891,6 +1023,8 @@ _(Agents designed to process and reason over multiple data types like images, te | Title | Venue | Date | Paper Link | Project Page | | :----------------------------------------------------------------------------------------------------------------------------------- | :------------------------------- | :------ | :---------------------------------------------------------------------- | :--------------------------------------------------------------------------------------------------------------------------------------------------------------- | +| **Towards Expert-level Medical AI for Real-time Video Consultations** | arXiv | 2026.08 | [Paper](https://arxiv.org/abs/2608.09861) | Not Available | +| **Guideline-as-Oracle: Zero-Annotation Training of an Ophthalmic Telephone Triage Agent** | arXiv | 2026.08 | [Paper](https://arxiv.org/abs/2608.04772) | Not Available | | **Towards Conversational Medical AI with Eyes, Ears and a Voice** | arXiv | 2026.05 | [Paper](https://arxiv.org/abs/2605.09272) | Not Available | | **SymptomAI: Toward a Conversational AI Agent for Everyday Symptom Assessment** | arXiv | 2026.05 | [Paper](https://arxiv.org/abs/2605.04012) | Not Available | | **ClinicBot: A Guideline-Grounded Clinical Chatbot with Prioritized Evidence RAG and Verifiable Citations** | arXiv | 2026.05 | [Paper](https://arxiv.org/abs/2605.00846) | Not Available | @@ -907,6 +1041,7 @@ _(Agents designed to process and reason over multiple data types like images, te | Title | Venue | Date | Paper Link | Project Page | | :-------------------------------------------------------------------------------------------------------------------------- | :-------------------------------- | :------ | :-------------------------------------------------------------------------- | :--------------------------------------------- | +| **Traj-Evolve: A Self-Evolving Multi-Agent System for Patient Trajectory Modeling in Lung Cancer Early Detection** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.02812) | Not Available | | **Detecting Clinical Discrepancies in Health Coaching Agents: A Dual-Stream Memory and Reconciliation Architecture** | arXiv | 2026.04 | [Paper](https://arxiv.org/abs/2604.27045) | Not Available | | **Agentic AI for Personalized Physiotherapy: A Multi-Agent Framework for Generative Video Training and Real-Time Pose Correction** | ICDH IEEE | 2026.04 | [Paper](https://arxiv.org/abs/2604.21154) | Not Available | | **Sense Less, Infer More: Agentic Multimodal Transformers for Edge Medical Intelligence** | arXiv | 2026.04 | [Paper](https://arxiv.org/abs/2604.10404) | Not Available | @@ -934,6 +1069,11 @@ _(Agents designed to process and reason over multiple data types like images, te | Title | Venue | Date | Paper Link | Project Page | | :--------------------------------------------------------------------------------------------------------- | :----------------------------- | :------ | :------------------------------------------------------------------- | :--------------------------------------------------------------------------------------------------------------------------------------------------- | +| **CASCADE: An Agentic Regulatory Network Framework for Patient-Data-Validated Downstream Perturbation Prediction** | arXiv | 2026.08 | [Paper](https://arxiv.org/abs/2608.05359) | Not Available | +| **An AI agent for treatment reasoning over a biomedical tool universe** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.28692) | ![Star](https://img.shields.io/github/stars/mims-harvard/ATHENA.svg?style=social&label=Star)
[GitHub](https://github.com/mims-harvard/ATHENA) | +| **BioInsight: Multi-Agent Orchestration for Interactive Biomedical Knowledge Discovery** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.20997) | Not Available | +| **DeepRoot: A KG-Coordinated Multi-Agent System for Therapeutic Reasoning over Historical Medical Texts** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.15931) | Not Available | +| **Beyond Prompt-Based Planning: MCP-Native Graph Planning-based Biomedical Agent System** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.04494) | Not Available | | **A Versatile AI Agent for Rare Disease Diagnosis and Risk Gene Prioritization (Hygieia)** | arXiv | 2026.05 | [Paper](https://arxiv.org/abs/2605.06226) | Not Available | | **FastOMOP: A Foundational Architecture for Reliable Agentic Real-World Evidence Generation on OMOP CDM data** | arXiv | 2026.04 | [Paper](https://arxiv.org/abs/2604.24572) | Not Available | | **Vibe Medicine: Redefining Biomedical Research Through Human-AI Co-Work** | arXiv | 2026.04 | [Paper](https://arxiv.org/abs/2604.23674) | Not Available | @@ -958,6 +1098,10 @@ _(Agents designed to process and reason over multiple data types like images, te | Title | Venue | Date | Paper Link | Project Page | | :-------------------------------------------------------------------------------------------------------------------------- | :----------------- | :------ | :----------------------------------------------------- | :----------------------------------------------------------------------------------------------------------------------------------------------------- | +| **From Single Chatbots to Governed Agent Ecosystems: An Agentic AI Pattern Catalogue and Orchestration Framework for Mission-Critical Hospital Information Management Systems** | arXiv | 2026.08 | [Paper](https://arxiv.org/abs/2608.07627) | Not Available | +| **From Siloed Algorithms to Compliance-First Agentic Platforms: A Multi-Layered Architecture for Hospital AI Systems** | arXiv | 2026.08 | [Paper](https://arxiv.org/abs/2608.06112) | Not Available | +| **Toward Trustworthy Large Language Model Agents in Healthcare** | arXiv | 2026.07 | [Paper](https://arxiv.org/abs/2607.05055) | ![Star](https://img.shields.io/github/stars/Hadi-Hsn/CareConnect.svg?style=social&label=Star)
[GitHub](https://github.com/Hadi-Hsn/CareConnect) | +| **Why Trust Your Agent? Empirical Security Gains from TRiSM-Guided Agentic Workflows in Healthcare** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.28666) | Not Available | | **CHI-Bench: Can AI Agents Automate End-to-End, Long-Horizon, Policy-Rich Healthcare Workflows?** | arXiv | 2026.05 | [Paper](https://arxiv.org/abs/2605.16679) | ![Star](https://img.shields.io/github/stars/actava-ai/chi-bench.svg?style=social&label=Star)
[GitHub](https://github.com/actava-ai/chi-bench)
[Project](https://actava.ai/benchmarks) | | **A Cross-Layered Multi-Drone Coordination for Medical Supply Delivery during Disaster Response Management** | arXiv | 2026.05 | [Paper](https://arxiv.org/abs/2605.09342) | Not Available | | **Research on Security Enhancement Methods for Adversarial Robust Large Language Model Intelligent Agents for Medical Decision-Making Tasks** | AINIT | 2026.05 | [Paper](https://arxiv.org/abs/2605.08257) | Not Available | @@ -996,6 +1140,25 @@ _(Agents designed to process and reason over multiple data types like images, te | Title | Venue | Date | Paper Link | Project Page | | :------------------------------------------------------------------------------------------------------------ | :---- | :------ | :----------------------------------------- | :------------------------------------------------------------------------------------------------------------------------------------------------------------------------- | +| **ELICITED: EHR-grounded Longitudinal Interactive Conversations for Information-seeking Triage Evaluation and Decision-making** | arXiv | 2026.08 | [Paper](https://arxiv.org/abs/2608.09024) | Not Available | +| **ClinLens: Towards Long-Horizon Coding Agents for Longitudinal Multimodal Clinical Data Science** | arXiv | 2026.07 | [Paper](https://arxiv.org/abs/2607.26155) | Not Available | +| **PatientAgentBench: A Benchmark Framework for Evaluating Patient-Facing Health AI Agents** | arXiv | 2026.07 | [Paper](https://arxiv.org/abs/2607.25485) | ![Star](https://img.shields.io/github/stars/amazon-science/PatientAgentBench.svg?style=social&label=Star)
[GitHub](https://github.com/amazon-science/PatientAgentBench) | +| **MedDDC-Eval: Diagnosis-Decoupled Evaluation of Multi-Turn Medical Consultation Agents** | arXiv | 2026.07 | [Paper](https://arxiv.org/abs/2607.18999) | Not Available | +| **Evaluating Frontier AI Agents as Autonomous Clinical Security Auditors** | arXiv | 2026.07 | [Paper](https://arxiv.org/abs/2607.13411) | Not Available | +| **LongMedBench: Benchmarking Medical Agents for Long-Horizon Clinical Decision-Making** | arXiv | 2026.07 | [Paper](https://arxiv.org/abs/2607.09322) | Not Available | +| **Evaluating Agentic Harness Systems for Autonomous Computational Pathology** | arXiv | 2026.07 | [Paper](https://arxiv.org/abs/2607.02598) | Not Available | +| **HealthAgentBench: A Unified Benchmark Suite of Realistic Agentic Healthcare Environments for Challenging Frontier AI Agents** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.31179) | ![Star](https://img.shields.io/github/stars/microsoft/HealthAgentBench.svg?style=social&label=Star)
[GitHub](https://github.com/microsoft/HealthAgentBench) | +| **MedEvoEval: Evaluating Continual Evolution of Doctor Agents through Simulated Clinical Episodes** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.28900) | Not Available | +| **EHR-Complex: Benchmarking Medical Agents for Complex Clinical Reasoning** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.23301) | Not Available | +| **OpenBioRQ: Unsolved Biomedical Research Questions for Agents** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.21959) | Not Available | +| **Are LLMs Ready to Assist Physicians? PhysAssistBench for Interactive Doctor-Patient-EHR Assistance** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.18613) | Not Available | +| **RubricsTree: Scalable and Evolving Open-Ended Evaluation of Personal Health Agents across Health Memory and Medical Skills** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.18203) | Not Available | +| **MedCTA: A Benchmark for Clinical Tool Agents** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.11702) | ![Star](https://img.shields.io/github/stars/IVUL-KAUST/MedCTA.svg?style=social&label=Star)
[GitHub](https://github.com/IVUL-KAUST/MedCTA) | +| **PSEBench: A Controllable and Verifiable Benchmark for Evaluating LLMs in Patient Safety Event Triage** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.05463) | Not Available | +| **MedCUA-Bench: A Screenshot-Only Benchmark for Clinical Computer-Use Agents** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.03203) | Not Available | +| **ClinicalMC: A Benchmark for Multi-Course Clinical Decision-Making with Large Language Models** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.03157) | Not Available | +| **ClinEnv: An Interactive Multi-Stage Long Horizon EHR Environment for Agents** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.02568) | Not Available | +| **AutoMedBench: Towards Medical AutoResearch with Agentic AI Models** | arXiv | 2026.06 | [Paper](https://arxiv.org/abs/2606.01961) | Not Available | | **CHI-Bench: Can AI Agents Automate End-to-End, Long-Horizon, Policy-Rich Healthcare Workflows?** | arXiv | 2026.05 | [Paper](https://arxiv.org/abs/2605.16679) | ![Star](https://img.shields.io/github/stars/actava-ai/chi-bench.svg?style=social&label=Star)
[GitHub](https://github.com/actava-ai/chi-bench) | | **MedMemoryBench: Benchmarking Agent Memory in Personalized Healthcare** | arXiv | 2026.05 | [Paper](https://arxiv.org/abs/2605.11814) | Not Available | | **ABRA: Agent Benchmark for Radiology Applications** | arXiv | 2026.05 | [Paper](https://arxiv.org/abs/2605.11224) | Not Available | @@ -1052,6 +1215,8 @@ _(Agents designed to process and reason over multiple data types like images, te | Title | Venue | Date | Paper Link | Project Page | | :---- | :---- | :--- | :--------- | :----------- | +| **Agentic AI in medicine: architectures, applications, evaluation, and challenges for clinical translation** | arXiv | 2026.07 | [Paper](https://arxiv.org/abs/2607.25489) | Not Available | +| **The Path to Self-Evolving Clinical Systems: Scaling Medical Agents from Assistance to Autonomy** | arXiv | 2026.07 | [Paper](https://arxiv.org/abs/2607.11175) | ![Star](https://img.shields.io/github/stars/zhcz328/Awesome-Medical-Agents.svg?style=social&label=Star)
[GitHub](https://github.com/zhcz328/Awesome-Medical-Agents) | | **Rethinking Health Agents: From Siloed AI to Collaborative Decision Mediators** | CHI Workshop | 2026 | [Paper](https://arxiv.org/abs/2603.24986) | Not Available | | **Six Interventions for the Responsible and Ethical Implementation of Medical AI Agents** | arXiv | 2026 | [Paper](https://arxiv.org/abs/2603.13743) | Not Available | | **A Comprehensive Survey of Agentic AI in Healthcare** | Authorea / TechRxiv | 2025 | [Paper](https://www.techrxiv.org/users/994756/articles/1355990-a-comprehensive-survey-of-agentic-ai-in-healthcare) | ![Star](https://img.shields.io/github/stars/AgenticHealthAI/Awesome-AI-Agents-for-Healthcare.svg?style=social&label=Star)
[GitHub](https://github.com/AgenticHealthAI/Awesome-AI-Agents-for-Healthcare) | @@ -1121,7 +1286,7 @@ Beyond academic research papers, several open-source projects and tools provide | **AI-Agents-for-Medical-Diagnostics** | LLM-based AI agents that analyze complex medical cases by integrating specialist insights | ![Star](https://img.shields.io/github/stars/ahmadvh/AI-Agents-for-Medical-Diagnostics.svg?style=social&label=Star)
[GitHub](https://github.com/ahmadvh/AI-Agents-for-Medical-Diagnostics) | | **HealthGPT (Stanford)** | Experimental iOS app for natural language interaction with Apple Health data | ![Star](https://img.shields.io/github/stars/StanfordBDHG/HealthGPT.svg?style=social&label=Star)
[GitHub](https://github.com/StanfordBDHG/HealthGPT) | | **DoctorGPT** | Offline-first LLM fine-tuned on medical dialogue data that can pass the US Medical Licensing Exam | ![Star](https://img.shields.io/github/stars/tmc/DoctorGPT.svg?style=social&label=Star)
[GitHub](https://github.com/tmc/DoctorGPT) | -| **MedSci Skills** | 32 open-source Claude Code skills for the full medical research lifecycle — anti-hallucination literature search (PubMed, Semantic Scholar, bioRxiv), meta-analysis pipeline (PROSPERO, PRISMA, QUADAS-2), reporting guideline audits (STROBE, PRISMA, STARD, CONSORT, TRIPOD+AI), statistical analysis in Python/R, publication-ready figures, study design review, grant proposals, peer review, and academic presentation prep. Three runnable end-to-end demos on public datasets. Built by a physician-researcher, MIT licensed | ![Star](https://img.shields.io/github/stars/Aperivue/medsci-skills.svg?style=social&label=Star)
[GitHub](https://github.com/Aperivue/medsci-skills) \| [Website](https://aperivue.com/skills) | +| **MedSci Skills** | 59 open-source agent skills for the full medical research lifecycle — anti-hallucination literature search (PubMed, Semantic Scholar, bioRxiv), meta-analysis pipeline (PROSPERO, PRISMA, QUADAS-2/QUADAS-3), reporting-guideline audits against 49 EQUATOR instruments (STROBE, PRISMA, STARD, CONSORT, TRIPOD+AI, CLAIM), statistical analysis in Python/R, publication-ready figures, study design review, grant proposals, peer review, and academic presentation prep. Five runnable end-to-end demos on public datasets, including external validation of a 3-D segmentation model across a modality shift. Runs in Claude Code, Codex, Cursor and GitHub Copilot. Built by a physician-researcher, MIT licensed | ![Star](https://img.shields.io/github/stars/Aperivue/medsci-skills.svg?style=social&label=Star)
[GitHub](https://github.com/Aperivue/medsci-skills) \| [Website](https://aperivue.com/skills) | | **Clinical AI Agent Skills** | Evidence-first agent rulebook for clinicians and medical AI builders using Codex, Claude Code, Cursor, GitHub Copilot, and Obsidian. Focuses on citation verification, research-vs-medical-advice boundaries, scoped file access, and safer clinical research workflows. | ![Star](https://img.shields.io/github/stars/2023Anita/clinical-ai-agent-skills.svg?style=social&label=Star)
[GitHub](https://github.com/2023Anita/clinical-ai-agent-skills) | | **Anthropic Healthcare Skills** | Official healthcare skills including FHIR developer tools, prior auth review, and clinical trial protocol generation | [GitHub](https://github.com/anthropics/healthcare) | | **Voice AI SDR Agent** | Production-ready autonomous AI phone agent for patient outreach, appointment scheduling, and healthcare communication using LangGraph and Twilio | ![Star](https://img.shields.io/github/stars/Rajathbharadwaj/voice-agent.svg?style=social&label=Star)
[GitHub](https://github.com/Rajathbharadwaj/voice-agent) | @@ -1144,7 +1309,6 @@ Beyond academic research papers, several open-source projects and tools provide | **Awesome Medical MCP Servers** | Curated collection of Medical MCP servers for healthcare data integration | ![Star](https://img.shields.io/github/stars/sunanhe/awesome-medical-mcp-servers.svg?style=social&label=Star)
[GitHub](https://github.com/sunanhe/awesome-medical-mcp-servers) | | **BGPT MCP** | Hosted MCP server for searching scientific papers with full-text experimental data extraction; covers biomedical, clinical, and life science studies; `search_papers` tool returns structured data (methods, results, sample sizes); 50 free searches | ![Star](https://img.shields.io/github/stars/connerlambden/bgpt-mcp.svg?style=social&label=Star)
[GitHub](https://github.com/connerlambden/bgpt-mcp) | | **Fulcra Context MCP** | Personal context MCP server providing unified access to biometric, sleep, activity, and calendar data for AI agents via the Fulcra Life API — enables healthcare AI applications to incorporate real-time patient-reported wellness data | ![Star](https://img.shields.io/github/stars/fulcradynamics/fulcra-context-mcp.svg?style=social&label=Star)
[GitHub](https://github.com/fulcradynamics/fulcra-context-mcp) \| [Python Client](https://github.com/fulcradynamics/fulcra-api-python) | - | **Genomic Agent Discovery** | Multi-agent MCP server for genomic analysis — specialized AI agents analyze raw DNA files across 12 databases (ClinVar, GWAS, AlphaMissense, CPIC, gnomAD, etc.) and coordinate findings through shared MCP tools. Privacy-first, runs 100% locally | ![Star](https://img.shields.io/github/stars/HelixGenomics/Genomic-Agent-Discovery.svg?style=social&label=Star)
[GitHub](https://github.com/HelixGenomics/Genomic-Agent-Discovery) | ## Healthcare RAG & Knowledge Systems @@ -1184,4 +1348,4 @@ To promote transparency and reproducibility, we provide the structured annotatio # Star History -[![Star History Chart](https://api.star-history.com/svg?repos=AgenticHealthAI/Awesome-AI-Agents-for-Healthcare&type=date&legend=top-left)](https://www.star-history.com/#AgenticHealthAI/Awesome-AI-Agents-for-Healthcare&type=date&legend=top-left) +[![Star History Chart](https://star-history.dera.page/svg?repos=AgenticHealthAI/Awesome-AI-Agents-for-Healthcare&type=date&legend=top-left)](https://star-history.dera.page/#AgenticHealthAI/Awesome-AI-Agents-for-Healthcare&type=date&legend=top-left)