AgentClinic: a multimodal agent benchmark to evaluate AI in simulated clinical environments
Fuente:
arXiv
Saved in:
| Main Authors: | Schmidgall, Samuel, Ziaei, Rojin, Harris, Carl, Reis, Eduardo, Jopling, Jeffrey, Moor, Michael |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Addressing cognitive bias in medical language models
by: Schmidgall, Samuel, et al.
Published: (2024)
by: Schmidgall, Samuel, et al.
Published: (2024)
Agent Laboratory: Using LLM Agents as Research Assistants
by: Schmidgall, Samuel, et al.
Published: (2025)
by: Schmidgall, Samuel, et al.
Published: (2025)
Beyond single-channel agentic benchmarking
by: Radpour, Nelu D.
Published: (2026)
by: Radpour, Nelu D.
Published: (2026)
PsychBench: A comprehensive and professional benchmark for evaluating the performance of LLM-assisted psychiatric clinical practice
by: Liu, Shuyu, et al.
Published: (2025)
by: Liu, Shuyu, et al.
Published: (2025)
A Combined Channel Approach for Decoding Intracranial EEG Signals: Enhancing Accuracy through Spatial Information Integration
by: Memar, Maryam Ostadsharif, et al.
Published: (2024)
by: Memar, Maryam Ostadsharif, et al.
Published: (2024)
CartoAgent: a multimodal large language model-powered multi-agent cartographic framework for map style transfer and evaluation
by: Wang, Chenglong, et al.
Published: (2025)
by: Wang, Chenglong, et al.
Published: (2025)
One Agent Too Many: User Perspectives on Approaches to Multi-agent Conversational AI
by: Clarke, Christopher, et al.
Published: (2024)
by: Clarke, Christopher, et al.
Published: (2024)
Simulating Cooperative Prosocial Behavior with Multi-Agent LLMs: Evidence and Mechanisms for AI Agents to Inform Policy Decisions
by: Sreedhar, Karthik, et al.
Published: (2025)
by: Sreedhar, Karthik, et al.
Published: (2025)
EyeAgent: An Agentic AI System for Multimodal Clinical Decision Support in Ophthalmology
by: Shi, Danli, et al.
Published: (2025)
by: Shi, Danli, et al.
Published: (2025)
Complementary Human-AI Clinical Reasoning in Ophthalmology
by: Sevgi, Mertcan, et al.
Published: (2025)
by: Sevgi, Mertcan, et al.
Published: (2025)
HIDAgent: A Toolkit Enabling "Personal Agents" on HID-Compatible Devices
by: Bigham, Jeffrey P.
Published: (2026)
by: Bigham, Jeffrey P.
Published: (2026)
From Interaction to Impact: Towards Safer AI Agents Through Understanding and Evaluating Mobile UI Operation Impacts
by: Zhang, Zhuohao Jerry, et al.
Published: (2024)
by: Zhang, Zhuohao Jerry, et al.
Published: (2024)
Evaluating multimodal emotion recognition in proactive conversational agents: A user study
by: Dragut, Adnana, et al.
Published: (2026)
by: Dragut, Adnana, et al.
Published: (2026)
Designing Human-AI Systems: Anthropomorphism and Framing Bias on Human-AI Collaboration
by: Olszewski, Samuel Aleksander Sánchez
Published: (2024)
by: Olszewski, Samuel Aleksander Sánchez
Published: (2024)
Multimodal Healthcare AI: Identifying and Designing Clinically Relevant Vision-Language Applications for Radiology
by: Yildirim, Nur, et al.
Published: (2024)
by: Yildirim, Nur, et al.
Published: (2024)
The Privacy Guardian Agent: Towards Trustworthy AI Privacy Agents
by: Freiberger, Vincent
Published: (2026)
by: Freiberger, Vincent
Published: (2026)
The Benefits of Prosociality towards AI Agents: Examining the Effects of Helping AI Agents on Human Well-Being
by: Zhu, Zicheng, et al.
Published: (2025)
by: Zhu, Zicheng, et al.
Published: (2025)
TAMA: A Human-AI Collaborative Thematic Analysis Framework Using Multi-Agent LLMs for Clinical Interviews
by: Xu, Huimin, et al.
Published: (2025)
by: Xu, Huimin, et al.
Published: (2025)
The Rise of AI Companions: Interaction with AI Companions and Psychological Well-being
by: Zhang, Yutong, et al.
Published: (2025)
by: Zhang, Yutong, et al.
Published: (2025)
Mapping the Design Space of User Experience for Computer Use Agents
by: Cheng, Ruijia, et al.
Published: (2026)
by: Cheng, Ruijia, et al.
Published: (2026)
ARise: an Augmented Reality Mobile Application to Improve Cultural Heritage Resilience
by: Urbanelli, Angelica, et al.
Published: (2025)
by: Urbanelli, Angelica, et al.
Published: (2025)
Voice-guided Orchestrated Intelligence for Clinical Evaluation (VOICE): A Voice AI Agent System for Prehospital Stroke Assessment
by: Acosta, Julian, et al.
Published: (2025)
by: Acosta, Julian, et al.
Published: (2025)
Optimising EEG decoding with refined sampling and multimodal feature integration
by: Akbarinia, Arash
Published: (2024)
by: Akbarinia, Arash
Published: (2024)
AnchoredAI: Contextual Anchoring of AI Comments Improves Writer Agency and Ownership
by: Lou, Martin, et al.
Published: (2025)
by: Lou, Martin, et al.
Published: (2025)
Leveraging Complementary AI Explanations to Mitigate Misunderstanding in XAI
by: Xuan, Yueqing, et al.
Published: (2025)
by: Xuan, Yueqing, et al.
Published: (2025)
Patterns for a New Generation: AI and Agents
by: Corneli, Joseph, et al.
Published: (2025)
by: Corneli, Joseph, et al.
Published: (2025)
"It depends on where AI is used": Players' attitude patterns and evaluative logics toward different AI applications in digital games
by: Hsu, Ting-Chen, et al.
Published: (2026)
by: Hsu, Ting-Chen, et al.
Published: (2026)
DiLLS: Interactive Diagnosis of LLM-based Multi-agent Systems via Layered Summary of Agent Behaviors
by: Sheng, Rui, et al.
Published: (2026)
by: Sheng, Rui, et al.
Published: (2026)
Augmenting Teamwork through AI Agents as Spatial Collaborators
by: Fernandez-Espinosa, Mariana, et al.
Published: (2025)
by: Fernandez-Espinosa, Mariana, et al.
Published: (2025)
Large-scale Evaluation of Notebook Checkpointing with AI Agents
by: Fang, Hanxi, et al.
Published: (2025)
by: Fang, Hanxi, et al.
Published: (2025)
Unremarkable to Remarkable AI Agent: Exploring Boundaries of Agent Intervention for Adults With and Without Cognitive Impairment
by: Chang, Mai Lee, et al.
Published: (2025)
by: Chang, Mai Lee, et al.
Published: (2025)
AgentClick: A Skill-Based Human-in-the-Loop Review Layer for Terminal AI Agents
by: Zhuang, Haomin, et al.
Published: (2026)
by: Zhuang, Haomin, et al.
Published: (2026)
Why Johnny Can't Use Agents: Industry Aspirations vs. User Realities with AI Agents
by: Shome, Pradyumna, et al.
Published: (2025)
by: Shome, Pradyumna, et al.
Published: (2025)
Towards spatial computing: recent advances in multimodal natural interaction for XR headsets
by: Wang, Zhimin, et al.
Published: (2025)
by: Wang, Zhimin, et al.
Published: (2025)
A learning health system in Neurorehabilitation as a foundation for multimodal patient representation
by: Weikert, Thomas, et al.
Published: (2026)
by: Weikert, Thomas, et al.
Published: (2026)
MICA: Medical Intelligent Conversational AgentHow to optimize medical teleconsultations for sports patients via a conversational agent?
by: Cervoni, Laurent, et al.
Published: (2024)
by: Cervoni, Laurent, et al.
Published: (2024)
AI Knows Best? The Paradox of Expertise, AI-Reliance, and Performance in Educational Tutoring Decision-Making Tasks
by: Chen, Eason, et al.
Published: (2025)
by: Chen, Eason, et al.
Published: (2025)
Changing human's impression of empathy from agent by verbalizing agent's position
by: Tsumura, Takahiro, et al.
Published: (2024)
by: Tsumura, Takahiro, et al.
Published: (2024)
Between Puppet and Actor: Reframing Authorship in this Age of AI Agents
by: Sun, Yuqian, et al.
Published: (2025)
by: Sun, Yuqian, et al.
Published: (2025)
Making AI Agents Evaluate Misleading Charts without Nudging
by: Panda, Swaroop
Published: (2026)
by: Panda, Swaroop
Published: (2026)
Similar Items
-
Addressing cognitive bias in medical language models
by: Schmidgall, Samuel, et al.
Published: (2024) -
Agent Laboratory: Using LLM Agents as Research Assistants
by: Schmidgall, Samuel, et al.
Published: (2025) -
Beyond single-channel agentic benchmarking
by: Radpour, Nelu D.
Published: (2026) -
PsychBench: A comprehensive and professional benchmark for evaluating the performance of LLM-assisted psychiatric clinical practice
by: Liu, Shuyu, et al.
Published: (2025) -
A Combined Channel Approach for Decoding Intracranial EEG Signals: Enhancing Accuracy through Spatial Information Integration
by: Memar, Maryam Ostadsharif, et al.
Published: (2024)