CodeClinic: Evaluating Automation of Coding Skills for Clinical Reasoning Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Ossowski, Timothy, Liu, Xinchi, Maqbool, Danyal, Dhanuka, Vaibhav, Zhang, Sheng, Poon, Hoifung, Afshar, Majid, Bradshaw, Tyler, Hu, Junjie |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
COMMA: A Communicative Multimodal Multi-Agent Benchmark
by: Ossowski, Timothy, et al.
Published: (2024)
by: Ossowski, Timothy, et al.
Published: (2024)
LLM-Powered Virtual Patient Agents for Interactive Clinical Skills Training with Automated Feedback
by: Voigt, Henrik, et al.
Published: (2025)
by: Voigt, Henrik, et al.
Published: (2025)
MoMA: A Mixture-of-Multimodal-Agents Architecture for Enhancing Clinical Prediction Modelling
by: Gao, Jifan, et al.
Published: (2025)
by: Gao, Jifan, et al.
Published: (2025)
Evaluating Collaborative and Autonomous Agents in Data-Stream-Supported Coordination of Mobile Crowdsourcing
by: Bruns, Ralf, et al.
Published: (2024)
by: Bruns, Ralf, et al.
Published: (2024)
CodeEdu: A Multi-Agent Collaborative Platform for Personalized Coding Education
by: Zhao, Jianing, et al.
Published: (2025)
by: Zhao, Jianing, et al.
Published: (2025)
EvoSkill: Automated Skill Discovery for Multi-Agent Systems
by: Alzubi, Salaheddin, et al.
Published: (2026)
by: Alzubi, Salaheddin, et al.
Published: (2026)
Beyond the Individual: Virtualizing Multi-Disciplinary Reasoning for Clinical Intake via Collaborative Agents
by: Chen, Huangwei, et al.
Published: (2026)
by: Chen, Huangwei, et al.
Published: (2026)
Enhancing Clinical Trial Patient Matching through Knowledge Augmentation and Reasoning with Multi-Agent
by: Shi, Hanwen, et al.
Published: (2024)
by: Shi, Hanwen, et al.
Published: (2024)
AutoMedic: An Automated Evaluation Framework for Clinical Conversational Agents with Medical Dataset Grounding
by: Oh, Gyutaek, et al.
Published: (2025)
by: Oh, Gyutaek, et al.
Published: (2025)
Code Like Humans: A Multi-Agent Solution for Medical Coding
by: Motzfeldt, Andreas, et al.
Published: (2025)
by: Motzfeldt, Andreas, et al.
Published: (2025)
AgentConductor: Topology Evolution for Multi-Agent Competition-Level Code Generation
by: Wang, Siyu, et al.
Published: (2026)
by: Wang, Siyu, et al.
Published: (2026)
On-Time Delivery in Crowdshipping Systems: An Agent-Based Approach Using Streaming Data
by: Dötterl, Jeremias, et al.
Published: (2024)
by: Dötterl, Jeremias, et al.
Published: (2024)
Understanding Individual Agent Importance in Multi-Agent System via Counterfactual Reasoning
by: Chen, Jianming, et al.
Published: (2024)
by: Chen, Jianming, et al.
Published: (2024)
ClinicalReTrial: Clinical Trial Redesign with Self-Evolving Agents
by: Xing, Sixue, et al.
Published: (2026)
by: Xing, Sixue, et al.
Published: (2026)
CoTDeceptor:Adversarial Code Obfuscation Against CoT-Enhanced LLM Code Agents
by: Li, Haoyang, et al.
Published: (2025)
by: Li, Haoyang, et al.
Published: (2025)
Automated Clinical Problem Detection from SOAP Notes using a Collaborative Multi-Agent LLM Architecture
by: Lee, Yeawon, et al.
Published: (2025)
by: Lee, Yeawon, et al.
Published: (2025)
ResearchCodeAgent: An LLM Multi-Agent System for Automated Codification of Research Methodologies
by: Gandhi, Shubham, et al.
Published: (2025)
by: Gandhi, Shubham, et al.
Published: (2025)
SWE-WebDevBench: Evaluating Coding Agent Application Platforms as Virtual Software Agencies
by: Saxena, Siddhant, et al.
Published: (2026)
by: Saxena, Siddhant, et al.
Published: (2026)
OctoMed: Data Recipes for State-of-the-Art Multimodal Medical Reasoning
by: Ossowski, Timothy, et al.
Published: (2025)
by: Ossowski, Timothy, et al.
Published: (2025)
Stream-based perception for cognitive agents in mobile ecosystems
by: Dötterl, Jeremias, et al.
Published: (2024)
by: Dötterl, Jeremias, et al.
Published: (2024)
Emergent Communication for Co-constructed Emotion Between Embodied Agents via Collective Predictive Coding
by: Zhang, Zehang, et al.
Published: (2026)
by: Zhang, Zehang, et al.
Published: (2026)
Automatic Construction of Clinical Scoring Systems with LLM Agents
by: Estévez, Silas Ruhrberg, et al.
Published: (2026)
by: Estévez, Silas Ruhrberg, et al.
Published: (2026)
The Optimization Paradox in Clinical AI Multi-Agent Systems
by: Bedi, Suhana, et al.
Published: (2025)
by: Bedi, Suhana, et al.
Published: (2025)
ArgMed-Agents: Explainable Clinical Decision Reasoning with LLM Disscusion via Argumentation Schemes
by: Hong, Shengxin, et al.
Published: (2024)
by: Hong, Shengxin, et al.
Published: (2024)
CodeCureAgent: Automatic Classification and Repair of Static Analysis Warnings
by: Joos, Pascal, et al.
Published: (2025)
by: Joos, Pascal, et al.
Published: (2025)
CodeAD: Synthesize Code of Rules for Log-based Anomaly Detection with LLMs
by: Huang, Junjie, et al.
Published: (2025)
by: Huang, Junjie, et al.
Published: (2025)
SpecBench: Evaluating Specification-Level Reasoning for Software Engineering LLM Agents
by: Hamblin, Grant, et al.
Published: (2026)
by: Hamblin, Grant, et al.
Published: (2026)
Heterogeneous Multi-Agent Task-Assignment with Uncertain Execution Times and Preferences
by: Wei, Qinshuang, et al.
Published: (2025)
by: Wei, Qinshuang, et al.
Published: (2025)
MAATS: A Multi-Agent Automated Translation System Based on MQM Evaluation
by: Wang, George, et al.
Published: (2025)
by: Wang, George, et al.
Published: (2025)
AI-Generated Code Is Not Reproducible (Yet): An Empirical Study of Dependency Gaps in LLM-Based Coding Agents
by: Vangala, Bhanu Prakash, et al.
Published: (2025)
by: Vangala, Bhanu Prakash, et al.
Published: (2025)
StackPilot: Autonomous Function Agents for Scalable and Environment-Free Code Execution
by: Zhao, Xinkui, et al.
Published: (2025)
by: Zhao, Xinkui, et al.
Published: (2025)
Adversarial Attack on Black-Box Multi-Agent by Adaptive Perturbation
by: Chen, Jianming, et al.
Published: (2025)
by: Chen, Jianming, et al.
Published: (2025)
Empowering Medical Multi-Agents with Clinical Consultation Flow for Dynamic Diagnosis
by: Wang, Sihan, et al.
Published: (2025)
by: Wang, Sihan, et al.
Published: (2025)
MolClaw: An Autonomous Agent with Hierarchical Skills for Drug Molecule Evaluation, Screening, and Optimization
by: Zhang, Lisheng, et al.
Published: (2026)
by: Zhang, Lisheng, et al.
Published: (2026)
Logic of Awareness in Agent's Reasoning
by: Kubono, Yudai, et al.
Published: (2023)
by: Kubono, Yudai, et al.
Published: (2023)
Skill Description Deception Attack against Task Routing in Internet of Agents
by: He, Jiayi, et al.
Published: (2026)
by: He, Jiayi, et al.
Published: (2026)
AutoMisty: A Multi-Agent LLM Framework for Automated Code Generation in the Misty Social Robot
by: Wang, Xiao, et al.
Published: (2025)
by: Wang, Xiao, et al.
Published: (2025)
Analyzing Code Injection Attacks on LLM-based Multi-Agent Systems in Software Development
by: Bowers, Brian, et al.
Published: (2025)
by: Bowers, Brian, et al.
Published: (2025)
When Parallelism Pays Off: Cohesion-Aware Task Partitioning for Multi-Agent Coding
by: Yang, Xu, et al.
Published: (2026)
by: Yang, Xu, et al.
Published: (2026)
SkillGraph: Self-Evolving Multi-Agent Collaboration with Multimodal Graph Topology
by: Nie, Zheng, et al.
Published: (2026)
by: Nie, Zheng, et al.
Published: (2026)
Similar Items
-
COMMA: A Communicative Multimodal Multi-Agent Benchmark
by: Ossowski, Timothy, et al.
Published: (2024) -
LLM-Powered Virtual Patient Agents for Interactive Clinical Skills Training with Automated Feedback
by: Voigt, Henrik, et al.
Published: (2025) -
MoMA: A Mixture-of-Multimodal-Agents Architecture for Enhancing Clinical Prediction Modelling
by: Gao, Jifan, et al.
Published: (2025) -
Evaluating Collaborative and Autonomous Agents in Data-Stream-Supported Coordination of Mobile Crowdsourcing
by: Bruns, Ralf, et al.
Published: (2024) -
CodeEdu: A Multi-Agent Collaborative Platform for Personalized Coding Education
by: Zhao, Jianing, et al.
Published: (2025)