Understanding and Optimizing Agentic Workflows via Shapley value
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Yingxuan, Huang, Bo, Qi, Siyuan, Feng, Chao, Hu, Haoyi, Zhu, Yuxuan, Hu, Jinbo, Zhao, Haoran, He, Ziyi, Liu, Xiao, Wen, Muning, Wang, Zongyu, Qiu, Lin, Cao, Xuezhi, Cai, Xunliang, Yu, Yong, Zhang, Weinan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Survey of AI Agent Protocols
von: Yang, Yingxuan, et al.
Veröffentlicht: (2025)
von: Yang, Yingxuan, et al.
Veröffentlicht: (2025)
DynaTree: Dynamic Agentic Retrieval Tree for Time-Sensitive News Retrieval
von: Qi, Siyuan, et al.
Veröffentlicht: (2026)
von: Qi, Siyuan, et al.
Veröffentlicht: (2026)
Why Not Act on What You Know? Unleashing Safety Potential of LLMs via Self-Aware Guard Enhancement
von: Ding, Peng, et al.
Veröffentlicht: (2025)
von: Ding, Peng, et al.
Veröffentlicht: (2025)
Understanding Agent Scaling in LLM-Based Multi-Agent Systems via Diversity
von: Yang, Yingxuan, et al.
Veröffentlicht: (2026)
von: Yang, Yingxuan, et al.
Veröffentlicht: (2026)
CoreCodeBench: Decoupling Code Intelligence via Fine-Grained Repository-Level Tasks
von: Fu, Lingyue, et al.
Veröffentlicht: (2025)
von: Fu, Lingyue, et al.
Veröffentlicht: (2025)
Position: Agentic AI System Is a Foreseeable Pathway to AGI
von: Liao, Junwei, et al.
Veröffentlicht: (2026)
von: Liao, Junwei, et al.
Veröffentlicht: (2026)
Friend or Foe: How LLMs' Safety Mind Gets Fooled by Intent Shift Attack
von: Ding, Peng, et al.
Veröffentlicht: (2025)
von: Ding, Peng, et al.
Veröffentlicht: (2025)
Automatically Benchmarking LLM Code Agents through Agent-Driven Annotation and Evaluation
von: Fu, Lingyue, et al.
Veröffentlicht: (2025)
von: Fu, Lingyue, et al.
Veröffentlicht: (2025)
P3: A Policy-Driven, Pace-Adaptive, and Diversity-Promoted Framework for data pruning in LLM Training
von: Yang, Yingxuan, et al.
Veröffentlicht: (2024)
von: Yang, Yingxuan, et al.
Veröffentlicht: (2024)
TRAD: Enhancing LLM Agents with Step-Wise Thought Retrieval and Aligned Decision
von: Zhou, Ruiwen, et al.
Veröffentlicht: (2024)
von: Zhou, Ruiwen, et al.
Veröffentlicht: (2024)
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese
von: Wang, Xihuai, et al.
Veröffentlicht: (2025)
von: Wang, Xihuai, et al.
Veröffentlicht: (2025)
SWE-Cycle: Benchmarking Code Agents across the Complete Issue Resolution Cycle
von: Guan, Hao, et al.
Veröffentlicht: (2026)
von: Guan, Hao, et al.
Veröffentlicht: (2026)
AMemGym: Interactive Memory Benchmarking for Assistants in Long-Horizon Conversations
von: Jiayang, Cheng, et al.
Veröffentlicht: (2026)
von: Jiayang, Cheng, et al.
Veröffentlicht: (2026)
CATArena: Evaluating Evolutionary Capabilities of Code Agents via Iterative Tournaments
von: Fu, Lingyue, et al.
Veröffentlicht: (2025)
von: Fu, Lingyue, et al.
Veröffentlicht: (2025)
MARFT: Multi-Agent Reinforcement Fine-Tuning
von: Liao, Junwei, et al.
Veröffentlicht: (2025)
von: Liao, Junwei, et al.
Veröffentlicht: (2025)
LARY: A Latent Action Representation Yielding Benchmark for Generalizable Vision-to-Action Alignment
von: Nie, Dujun, et al.
Veröffentlicht: (2026)
von: Nie, Dujun, et al.
Veröffentlicht: (2026)
Benchmarking Agentic Workflow Generation
von: Qiao, Shuofei, et al.
Veröffentlicht: (2024)
von: Qiao, Shuofei, et al.
Veröffentlicht: (2024)
AgentNet: Decentralized Evolutionary Coordination for LLM-based Multi-Agent Systems
von: Yang, Yingxuan, et al.
Veröffentlicht: (2025)
von: Yang, Yingxuan, et al.
Veröffentlicht: (2025)
Sherlock: Reliable and Efficient Agentic Workflow Execution
von: Ro, Yeonju, et al.
Veröffentlicht: (2025)
von: Ro, Yeonju, et al.
Veröffentlicht: (2025)
GradPower: Powering Gradients for Faster Language Model Pre-Training
von: Wang, Jinbo, et al.
Veröffentlicht: (2025)
von: Wang, Jinbo, et al.
Veröffentlicht: (2025)
Murakkab: Resource-Efficient Agentic Workflow Orchestration in Cloud Platforms
von: Chaudhry, Gohar Irfan, et al.
Veröffentlicht: (2025)
von: Chaudhry, Gohar Irfan, et al.
Veröffentlicht: (2025)
Instance-level Randomization: Toward More Stable LLM Evaluations
von: Li, Yiyang, et al.
Veröffentlicht: (2025)
von: Li, Yiyang, et al.
Veröffentlicht: (2025)
EnvX: Agentize Everything with Agentic AI
von: Chen, Linyao, et al.
Veröffentlicht: (2025)
von: Chen, Linyao, et al.
Veröffentlicht: (2025)
Autonomous Goal Detection and Cessation in Reinforcement Learning: A Case Study on Source Term Estimation
von: Shi, Yiwei, et al.
Veröffentlicht: (2024)
von: Shi, Yiwei, et al.
Veröffentlicht: (2024)
GLOW: Graph-Language Co-Reasoning for Agentic Workflow Performance Prediction
von: Guan, Wei, et al.
Veröffentlicht: (2025)
von: Guan, Wei, et al.
Veröffentlicht: (2025)
Reinforcing Language Agents via Policy Optimization with Action Decomposition
von: Wen, Muning, et al.
Veröffentlicht: (2024)
von: Wen, Muning, et al.
Veröffentlicht: (2024)
Agent Exchange: Shaping the Future of AI Agent Economics
von: Yang, Yingxuan, et al.
Veröffentlicht: (2025)
von: Yang, Yingxuan, et al.
Veröffentlicht: (2025)
UNO-Bench: A Unified Benchmark for Exploring the Compositional Law Between Uni-modal and Omni-modal in Omni Models
von: Chen, Chen, et al.
Veröffentlicht: (2025)
von: Chen, Chen, et al.
Veröffentlicht: (2025)
SALM: Spatial Audio Language Model with Structured Embeddings for Understanding and Editing
von: Hu, Jinbo, et al.
Veröffentlicht: (2025)
von: Hu, Jinbo, et al.
Veröffentlicht: (2025)
Leveraging Dual Process Theory in Language Agent Framework for Real-time Simultaneous Human-AI Collaboration
von: Zhang, Shao, et al.
Veröffentlicht: (2025)
von: Zhang, Shao, et al.
Veröffentlicht: (2025)
Fast Catch-Up, Late Switching: Optimal Batch Size Scheduling via Functional Scaling Laws
von: Wang, Jinbo, et al.
Veröffentlicht: (2026)
von: Wang, Jinbo, et al.
Veröffentlicht: (2026)
Holos: A Web-Scale LLM-Based Multi-Agent System for the Agentic Web
von: Nie, Xiaohang, et al.
Veröffentlicht: (2026)
von: Nie, Xiaohang, et al.
Veröffentlicht: (2026)
Agentic Web: Weaving the Next Web with AI Agents
von: Yang, Yingxuan, et al.
Veröffentlicht: (2025)
von: Yang, Yingxuan, et al.
Veröffentlicht: (2025)
Agentic Workflow for Education: Concepts and Applications
von: Jiang, Yuan-Hao, et al.
Veröffentlicht: (2025)
von: Jiang, Yuan-Hao, et al.
Veröffentlicht: (2025)
EvoFlow: Evolving Diverse Agentic Workflows On The Fly
von: Zhang, Guibin, et al.
Veröffentlicht: (2025)
von: Zhang, Guibin, et al.
Veröffentlicht: (2025)
Length Desensitization in Direct Preference Optimization
von: Liu, Wei, et al.
Veröffentlicht: (2024)
von: Liu, Wei, et al.
Veröffentlicht: (2024)
AgentEscapeBench: Evaluating Out-of-Domain Tool-Grounded Reasoning in LLM Agents
von: Guo, Zhengkang, et al.
Veröffentlicht: (2026)
von: Guo, Zhengkang, et al.
Veröffentlicht: (2026)
Entropy-Regularized Token-Level Policy Optimization for Language Agent Reinforcement
von: Wen, Muning, et al.
Veröffentlicht: (2024)
von: Wen, Muning, et al.
Veröffentlicht: (2024)
Helpful Agent Meets Deceptive Judge: Understanding Vulnerabilities in Agentic Workflows
von: Ming, Yifei, et al.
Veröffentlicht: (2025)
von: Ming, Yifei, et al.
Veröffentlicht: (2025)
LiveFMBench: Unveiling the Power and Limits of Agentic Workflows in Specification Generation
von: Xu, Dong, et al.
Veröffentlicht: (2026)
von: Xu, Dong, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
A Survey of AI Agent Protocols
von: Yang, Yingxuan, et al.
Veröffentlicht: (2025) -
DynaTree: Dynamic Agentic Retrieval Tree for Time-Sensitive News Retrieval
von: Qi, Siyuan, et al.
Veröffentlicht: (2026) -
Why Not Act on What You Know? Unleashing Safety Potential of LLMs via Self-Aware Guard Enhancement
von: Ding, Peng, et al.
Veröffentlicht: (2025) -
Understanding Agent Scaling in LLM-Based Multi-Agent Systems via Diversity
von: Yang, Yingxuan, et al.
Veröffentlicht: (2026) -
CoreCodeBench: Decoupling Code Intelligence via Fine-Grained Repository-Level Tasks
von: Fu, Lingyue, et al.
Veröffentlicht: (2025)