Guardado en:
| Autores principales: | Cotton, R. James, Leonard, Thomas |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2602.06975 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
BiomechGPT: Towards a Biomechanically Fluent Multimodal Foundation Model for Clinically Relevant Motion Tasks
por: Yang, Ruize, et al.
Publicado: (2025)
por: Yang, Ruize, et al.
Publicado: (2025)
Code as Agent Harness
por: Ning, Xuying, et al.
Publicado: (2026)
por: Ning, Xuying, et al.
Publicado: (2026)
DA-Code: Agent Data Science Code Generation Benchmark for Large Language Models
por: Huang, Yiming, et al.
Publicado: (2024)
por: Huang, Yiming, et al.
Publicado: (2024)
Paper2Agent: Reimagining Research Papers As Interactive and Reliable AI Agents
por: Miao, Jiacheng, et al.
Publicado: (2025)
por: Miao, Jiacheng, et al.
Publicado: (2025)
Agent-Testing Agent: A Meta-Agent for Automated Testing and Evaluation of Conversational AI Agents
por: Komoravolu, Sameer, et al.
Publicado: (2025)
por: Komoravolu, Sameer, et al.
Publicado: (2025)
Impatient Users Confuse AI Agents: High-fidelity Simulations of Human Traits for Testing Agents
por: He, Muyu, et al.
Publicado: (2025)
por: He, Muyu, et al.
Publicado: (2025)
MapCoder: Multi-Agent Code Generation for Competitive Problem Solving
por: Islam, Md. Ashraful, et al.
Publicado: (2024)
por: Islam, Md. Ashraful, et al.
Publicado: (2024)
AI Knowledge Assist: An Automated Approach for the Creation of Knowledge Bases for Conversational AI Agents
por: Laskar, Md Tahmid Rahman, et al.
Publicado: (2025)
por: Laskar, Md Tahmid Rahman, et al.
Publicado: (2025)
Monocular Biomechanical Tracking of Fingers with Inverse Kinematics to Foundation Models
por: Cotton, R. James, et al.
Publicado: (2026)
por: Cotton, R. James, et al.
Publicado: (2026)
AutoAgent: A Fully-Automated and Zero-Code Framework for LLM Agents
por: Tang, Jiabin, et al.
Publicado: (2025)
por: Tang, Jiabin, et al.
Publicado: (2025)
Holistic Agent Leaderboard: The Missing Infrastructure for AI Agent Evaluation
por: Kapoor, Sayash, et al.
Publicado: (2025)
por: Kapoor, Sayash, et al.
Publicado: (2025)
Coding Agents are Effective Long-Context Processors
por: Cao, Weili, et al.
Publicado: (2026)
por: Cao, Weili, et al.
Publicado: (2026)
Memory in the Age of AI Agents
por: Hu, Yuyang, et al.
Publicado: (2025)
por: Hu, Yuyang, et al.
Publicado: (2025)
MOSS: Enabling Code-Driven Evolution and Context Management for AI Agents
por: Zhu, Ming, et al.
Publicado: (2024)
por: Zhu, Ming, et al.
Publicado: (2024)
LocAgent: Graph-Guided LLM Agents for Code Localization
por: Chen, Zhaoling, et al.
Publicado: (2025)
por: Chen, Zhaoling, et al.
Publicado: (2025)
InfiAgent-DABench: Evaluating Agents on Data Analysis Tasks
por: Hu, Xueyu, et al.
Publicado: (2024)
por: Hu, Xueyu, et al.
Publicado: (2024)
CodeScout: An Effective Recipe for Reinforcement Learning of Code Search Agents
por: Sutawika, Lintang, et al.
Publicado: (2026)
por: Sutawika, Lintang, et al.
Publicado: (2026)
Executable Code Actions Elicit Better LLM Agents
por: Wang, Xingyao, et al.
Publicado: (2024)
por: Wang, Xingyao, et al.
Publicado: (2024)
DocAgent: A Multi-Agent System for Automated Code Documentation Generation
por: Yang, Dayu, et al.
Publicado: (2025)
por: Yang, Dayu, et al.
Publicado: (2025)
CR-Bench: Evaluating the Real-World Utility of AI Code Review Agents
por: Pereira, Kristen, et al.
Publicado: (2026)
por: Pereira, Kristen, et al.
Publicado: (2026)
MARS$^2$: Scaling Multi-Agent Tree Search via Reinforcement Learning for Code Generation
por: Li, Pengfei, et al.
Publicado: (2026)
por: Li, Pengfei, et al.
Publicado: (2026)
Multi-Agent System for AI-Assisted Extraction of Narrative Arcs in TV Series
por: Balestri, Roberto, et al.
Publicado: (2025)
por: Balestri, Roberto, et al.
Publicado: (2025)
Questionnaire Responses Do not Capture the Safety of AI Agents
por: Hellrigel-Holderbaum, Max, et al.
Publicado: (2026)
por: Hellrigel-Holderbaum, Max, et al.
Publicado: (2026)
Distilling LLM Agent into Small Models with Retrieval and Code Tools
por: Kang, Minki, et al.
Publicado: (2025)
por: Kang, Minki, et al.
Publicado: (2025)
CACA Agent: Capability Collaboration based AI Agent
por: Xu, Peng, et al.
Publicado: (2024)
por: Xu, Peng, et al.
Publicado: (2024)
Driving Generative Agents With Their Personality
por: Klinkert, Lawrence J., et al.
Publicado: (2024)
por: Klinkert, Lawrence J., et al.
Publicado: (2024)
Holistic Evaluation and Failure Diagnosis of AI Agents
por: Madvil, Netta, et al.
Publicado: (2026)
por: Madvil, Netta, et al.
Publicado: (2026)
Conformity and Social Impact on AI Agents
por: Bellina, Alessandro, et al.
Publicado: (2026)
por: Bellina, Alessandro, et al.
Publicado: (2026)
Self-Explanation in Social AI Agents
por: Basappa, Rhea, et al.
Publicado: (2025)
por: Basappa, Rhea, et al.
Publicado: (2025)
A Survey on Code Generation with LLM-based Agents
por: Dong, Yihong, et al.
Publicado: (2025)
por: Dong, Yihong, et al.
Publicado: (2025)
API-Assisted Code Generation for Question Answering on Varied Table Structures
por: Cao, Yihan, et al.
Publicado: (2023)
por: Cao, Yihan, et al.
Publicado: (2023)
Breaking the Data Barrier -- Building GUI Agents Through Task Generalization
por: Zhang, Junlei, et al.
Publicado: (2025)
por: Zhang, Junlei, et al.
Publicado: (2025)
Policy-as-Prompt: Turning AI Governance Rules into Guardrails for AI Agents
por: Kholkar, Gauri, et al.
Publicado: (2025)
por: Kholkar, Gauri, et al.
Publicado: (2025)
Context Pruning for Coding Agents via Multi-Rubric Latent Reasoning
por: Wang, Jingjing, et al.
Publicado: (2026)
por: Wang, Jingjing, et al.
Publicado: (2026)
CATArena: Evaluating Evolutionary Capabilities of Code Agents via Iterative Tournaments
por: Fu, Lingyue, et al.
Publicado: (2025)
por: Fu, Lingyue, et al.
Publicado: (2025)
Code Broker: A Multi-Agent System for Automated Code Quality Assessment
por: Attrah, Samer
Publicado: (2026)
por: Attrah, Samer
Publicado: (2026)
Preserving Cultural Identity with Context-Aware Translation Through Multi-Agent AI Systems
por: Anik, Mahfuz Ahmed, et al.
Publicado: (2025)
por: Anik, Mahfuz Ahmed, et al.
Publicado: (2025)
AgentRefine: Enhancing Agent Generalization through Refinement Tuning
por: Fu, Dayuan, et al.
Publicado: (2025)
por: Fu, Dayuan, et al.
Publicado: (2025)
AI Planning Framework for LLM-Based Web Agents
por: Shahnovsky, Orit, et al.
Publicado: (2026)
por: Shahnovsky, Orit, et al.
Publicado: (2026)
Characteristic AI Agents via Large Language Models
por: Wang, Xi, et al.
Publicado: (2024)
por: Wang, Xi, et al.
Publicado: (2024)
Ejemplares similares
-
BiomechGPT: Towards a Biomechanically Fluent Multimodal Foundation Model for Clinically Relevant Motion Tasks
por: Yang, Ruize, et al.
Publicado: (2025) -
Code as Agent Harness
por: Ning, Xuying, et al.
Publicado: (2026) -
DA-Code: Agent Data Science Code Generation Benchmark for Large Language Models
por: Huang, Yiming, et al.
Publicado: (2024) -
Paper2Agent: Reimagining Research Papers As Interactive and Reliable AI Agents
por: Miao, Jiacheng, et al.
Publicado: (2025) -
Agent-Testing Agent: A Meta-Agent for Automated Testing and Evaluation of Conversational AI Agents
por: Komoravolu, Sameer, et al.
Publicado: (2025)