The End of Reward Engineering: How LLMs Are Redefining Multi-Agent Coordination
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Su, Haoran, Sun, Yandong, Yu, Congjia |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Spatiotemporal Decision Transformer for Traffic Coordination
von: Su, Haoran, et al.
Veröffentlicht: (2026)
von: Su, Haoran, et al.
Veröffentlicht: (2026)
Emergency Preemption Without Online Exploration: A Decision Transformer Approach
von: Su, Haoran, et al.
Veröffentlicht: (2026)
von: Su, Haoran, et al.
Veröffentlicht: (2026)
Multi-Agent Coordination Adaptation via Structure-Guided Orchestration
von: Li, Haoran, et al.
Veröffentlicht: (2026)
von: Li, Haoran, et al.
Veröffentlicht: (2026)
Alternating Target-Path Planning for Scalable Multi-Agent Coordination
von: Kumagai, Yu, et al.
Veröffentlicht: (2026)
von: Kumagai, Yu, et al.
Veröffentlicht: (2026)
RewardHackingAgents: Benchmarking Evaluation Integrity for LLM ML-Engineering Agents
von: Atinafu, Yonas, et al.
Veröffentlicht: (2026)
von: Atinafu, Yonas, et al.
Veröffentlicht: (2026)
SA-IQA: Redefining Image Quality Assessment for Spatial Aesthetics with Multi-Dimensional Rewards
von: Gao, Yuan, et al.
Veröffentlicht: (2025)
von: Gao, Yuan, et al.
Veröffentlicht: (2025)
CausalAgent: A Conversational Multi-Agent System for End-to-End Causal Inference
von: Zhu, Jiawei, et al.
Veröffentlicht: (2026)
von: Zhu, Jiawei, et al.
Veröffentlicht: (2026)
CalBench: Evaluating Coordination-Privacy Trade-offs in Multi-Agent LLMs
von: Zou, Chelsea, et al.
Veröffentlicht: (2026)
von: Zou, Chelsea, et al.
Veröffentlicht: (2026)
NORA: A Harness-Engineered Autonomous Research Agent for End-to-End Spatial Data Science
von: Zhou, Bing, et al.
Veröffentlicht: (2026)
von: Zhou, Bing, et al.
Veröffentlicht: (2026)
Introspection of Thought Helps AI Agents
von: Sun, Haoran, et al.
Veröffentlicht: (2025)
von: Sun, Haoran, et al.
Veröffentlicht: (2025)
Facilitating Emergency Vehicle Passage in Congested Urban Areas Using Multi-agent Deep Reinforcement Learning
von: Su, Haoran
Veröffentlicht: (2025)
von: Su, Haoran
Veröffentlicht: (2025)
Multi-Agent Coordination across Diverse Applications: A Survey
von: Sun, Lijun, et al.
Veröffentlicht: (2025)
von: Sun, Lijun, et al.
Veröffentlicht: (2025)
How LLMs Follow Instructions: Skillful Coordination, Not a Universal Mechanism
von: Rocchetti, Elisabetta, et al.
Veröffentlicht: (2026)
von: Rocchetti, Elisabetta, et al.
Veröffentlicht: (2026)
Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL
von: Li, Weizhen, et al.
Veröffentlicht: (2025)
von: Li, Weizhen, et al.
Veröffentlicht: (2025)
Cooperative Reward Shaping for Multi-Agent Pathfinding
von: Song, Zhenyu, et al.
Veröffentlicht: (2024)
von: Song, Zhenyu, et al.
Veröffentlicht: (2024)
MetaAgent-X : Breaking the Ceiling of Automatic Multi-Agent Systems via End-to-End Reinforcement Learning
von: Zhang, Yaolun, et al.
Veröffentlicht: (2026)
von: Zhang, Yaolun, et al.
Veröffentlicht: (2026)
Agent-RLVR: Training Software Engineering Agents via Guidance and Environment Rewards
von: Da, Jeff, et al.
Veröffentlicht: (2025)
von: Da, Jeff, et al.
Veröffentlicht: (2025)
Confidence as a Reward: Transforming LLMs into Reward Models
von: Du, He, et al.
Veröffentlicht: (2025)
von: Du, He, et al.
Veröffentlicht: (2025)
Resilient Multi-Agent Negotiation for Medical Supply Chains:Integrating LLMs and Blockchain for Transparent Coordination
von: ALMutairi, Mariam, et al.
Veröffentlicht: (2025)
von: ALMutairi, Mariam, et al.
Veröffentlicht: (2025)
Data-Efficient Multi-Agent Spatial Planning with LLMs
von: Su, Huangyuan, et al.
Veröffentlicht: (2025)
von: Su, Huangyuan, et al.
Veröffentlicht: (2025)
Swarm Skills: A Portable, Self-Evolving Multi-Agent System Specification for Coordination Engineering
von: Zhang, Xinyu, et al.
Veröffentlicht: (2026)
von: Zhang, Xinyu, et al.
Veröffentlicht: (2026)
Omni-Thinker: Scaling Multi-Task RL in LLMs with Hybrid Reward and Task Scheduling
von: Li, Derek, et al.
Veröffentlicht: (2025)
von: Li, Derek, et al.
Veröffentlicht: (2025)
GOV-REK: Governed Reward Engineering Kernels for Designing Robust Multi-Agent Reinforcement Learning Systems
von: Rana, Ashish, et al.
Veröffentlicht: (2024)
von: Rana, Ashish, et al.
Veröffentlicht: (2024)
Coordination Graphs for Constrained Multi-Agent Reinforcement Learning
von: Amaya-Corredor, Santiago, et al.
Veröffentlicht: (2026)
von: Amaya-Corredor, Santiago, et al.
Veröffentlicht: (2026)
AGORA: Adapter-Grounded Observation-Action Retention for Inference-Free Prompt Compression in LLM Agents
von: Zhang, Haoran, et al.
Veröffentlicht: (2026)
von: Zhang, Haoran, et al.
Veröffentlicht: (2026)
FROGENT: An End-to-End Full-process Drug Design Multi-Agent System
von: Pan, Qihua, et al.
Veröffentlicht: (2025)
von: Pan, Qihua, et al.
Veröffentlicht: (2025)
Multi-Agent Coordinated Rename Refactoring
von: Bellur, Abhiram, et al.
Veröffentlicht: (2026)
von: Bellur, Abhiram, et al.
Veröffentlicht: (2026)
Multi-Agent Reinforcement Learning with a Hierarchy of Reward Machines
von: Zheng, Xuejing, et al.
Veröffentlicht: (2024)
von: Zheng, Xuejing, et al.
Veröffentlicht: (2024)
How to Build AI Agents by Augmenting LLMs with Codified Human Expert Domain Knowledge? A Software Engineering Framework
von: uulu, Choro Ulan, et al.
Veröffentlicht: (2026)
von: uulu, Choro Ulan, et al.
Veröffentlicht: (2026)
PythonSaga: Redefining the Benchmark to Evaluate Code Generating LLMs
von: Yadav, Ankit, et al.
Veröffentlicht: (2024)
von: Yadav, Ankit, et al.
Veröffentlicht: (2024)
ARMS: Automatic Reward Shaping for Sparse-Reward Multi-Agent Reinforcement Learning
von: Abboud, Elie, et al.
Veröffentlicht: (2026)
von: Abboud, Elie, et al.
Veröffentlicht: (2026)
Learning to Communicate: Toward End-to-End Optimization of Multi-Agent Language Systems
von: Yu, Ye, et al.
Veröffentlicht: (2026)
von: Yu, Ye, et al.
Veröffentlicht: (2026)
Prompting Multi-Modal Tokens to Enhance End-to-End Autonomous Driving Imitation Learning with LLMs
von: Duan, Yiqun, et al.
Veröffentlicht: (2024)
von: Duan, Yiqun, et al.
Veröffentlicht: (2024)
From Scores to Preferences: Redefining MOS Benchmarking for Speech Quality Reward Modeling
von: Cao, Yifei, et al.
Veröffentlicht: (2025)
von: Cao, Yifei, et al.
Veröffentlicht: (2025)
How Far Are We on the Decision-Making of LLMs? Evaluating LLMs' Gaming Ability in Multi-Agent Environments
von: Huang, Jen-tse, et al.
Veröffentlicht: (2024)
von: Huang, Jen-tse, et al.
Veröffentlicht: (2024)
SWE-Dev: Building Software Engineering Agents with Training and Inference Scaling
von: Wang, Haoran, et al.
Veröffentlicht: (2025)
von: Wang, Haoran, et al.
Veröffentlicht: (2025)
Multi-Agent Collaborative Reward Design for Enhancing Reasoning in Reinforcement Learning
von: Yang, Pei, et al.
Veröffentlicht: (2025)
von: Yang, Pei, et al.
Veröffentlicht: (2025)
Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies
von: Li, Zhuoran, et al.
Veröffentlicht: (2026)
von: Li, Zhuoran, et al.
Veröffentlicht: (2026)
Adaptive Theory of Mind for LLM-based Multi-Agent Coordination
von: Mu, Chunjiang, et al.
Veröffentlicht: (2026)
von: Mu, Chunjiang, et al.
Veröffentlicht: (2026)
Enhancing Presentation Slide Generation by LLMs with a Multi-Staged End-to-End Approach
von: Bandyopadhyay, Sambaran, et al.
Veröffentlicht: (2024)
von: Bandyopadhyay, Sambaran, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Spatiotemporal Decision Transformer for Traffic Coordination
von: Su, Haoran, et al.
Veröffentlicht: (2026) -
Emergency Preemption Without Online Exploration: A Decision Transformer Approach
von: Su, Haoran, et al.
Veröffentlicht: (2026) -
Multi-Agent Coordination Adaptation via Structure-Guided Orchestration
von: Li, Haoran, et al.
Veröffentlicht: (2026) -
Alternating Target-Path Planning for Scalable Multi-Agent Coordination
von: Kumagai, Yu, et al.
Veröffentlicht: (2026) -
RewardHackingAgents: Benchmarking Evaluation Integrity for LLM ML-Engineering Agents
von: Atinafu, Yonas, et al.
Veröffentlicht: (2026)