One Step is Enough: Multi-Agent Reinforcement Learning based on One-Step Policy Optimization for Order Dispatch on Ride-Sharing Platforms
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhao, Zijian, Li, Sen |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
BMG-Q: Localized Bipartite Match Graph Attention Q-Learning for Ride-Pooling Order Dispatch
por: Hu, Yulong, et al.
Publicado: (2025)
por: Hu, Yulong, et al.
Publicado: (2025)
Triple-BERT: Do We Really Need MARL for Order Dispatch on Ride-Sharing Platforms?
por: Zhao, Zijian, et al.
Publicado: (2025)
por: Zhao, Zijian, et al.
Publicado: (2025)
LCGuard: Latent Communication Guard for Safe KV Sharing in Multi-Agent Systems
por: Asif, Sadia, et al.
Publicado: (2026)
por: Asif, Sadia, et al.
Publicado: (2026)
Autonomous Traffic Signal Optimization Using Digital Twin and Agentic AI for Real-Time Decision-Making
por: Jan, Salman, et al.
Publicado: (2026)
por: Jan, Salman, et al.
Publicado: (2026)
ClinEnv: An Interactive Multi-Stage Long Horizon EHR Environment for Agents
por: Lu, Yuxing, et al.
Publicado: (2026)
por: Lu, Yuxing, et al.
Publicado: (2026)
Content Caching-Assisted Vehicular Edge Computing Using Multi-Agent Graph Attention Reinforcement Learning
por: Shen, Jinjin, et al.
Publicado: (2024)
por: Shen, Jinjin, et al.
Publicado: (2024)
Scale-Plan: Scalable Language-Enabled Task Planning for Heterogeneous Multi-Robot Teams
por: Gupta, Piyush, et al.
Publicado: (2026)
por: Gupta, Piyush, et al.
Publicado: (2026)
Synergistic Simulations: Multi-Agent Problem Solving with Large Language Models
por: Sprigler, Asher, et al.
Publicado: (2024)
por: Sprigler, Asher, et al.
Publicado: (2024)
Training-Free Agentic AI: Probabilistic Control and Coordination in Multi-Agent LLM Systems
por: Hosseini, Mohammad Parsa, et al.
Publicado: (2026)
por: Hosseini, Mohammad Parsa, et al.
Publicado: (2026)
MI9: An Integrated Runtime Governance Framework for Agentic AI
por: Wang, Charles L., et al.
Publicado: (2025)
por: Wang, Charles L., et al.
Publicado: (2025)
Towards a HIPAA Compliant Agentic AI System in Healthcare
por: Neupane, Subash, et al.
Publicado: (2025)
por: Neupane, Subash, et al.
Publicado: (2025)
TruthTensor: Evaluating LLMs through Human Imitation on Prediction Market under Drift and Holistic Reasoning
por: Shahabi, Shirin, et al.
Publicado: (2026)
por: Shahabi, Shirin, et al.
Publicado: (2026)
LLM experiments with simulation: Large Language Model Multi-Agent System for Simulation Model Parametrization in Digital Twins
por: Xia, Yuchen, et al.
Publicado: (2024)
por: Xia, Yuchen, et al.
Publicado: (2024)
KGLAMP: Knowledge Graph-guided Language model for Adaptive Multi-robot Planning and Replanning
por: Shek, Chak Lam, et al.
Publicado: (2026)
por: Shek, Chak Lam, et al.
Publicado: (2026)
SentinelAI: A Multi-Agent Framework for Structuring and Linking NG9-1-1 Emergency Incident Data
por: Ho, Kliment, et al.
Publicado: (2026)
por: Ho, Kliment, et al.
Publicado: (2026)
On the Ethical Considerations of Generative Agents
por: Diamond, N'yoma, et al.
Publicado: (2024)
por: Diamond, N'yoma, et al.
Publicado: (2024)
Demystifying AI Agents: The Final Generation of Intelligence
por: McNamara, Kevin J, et al.
Publicado: (2025)
por: McNamara, Kevin J, et al.
Publicado: (2025)
Interacting Large Language Model Agents. Interpretable Models and Social Learning
por: Jain, Adit, et al.
Publicado: (2024)
por: Jain, Adit, et al.
Publicado: (2024)
Multi-agent Systems for Misinformation Lifecycle : Detection, Correction And Source Identification
por: Gautam, Aditya
Publicado: (2025)
por: Gautam, Aditya
Publicado: (2025)
Urban Air Mobility as a System of Systems: An LLM-Enhanced Holonic Approach
por: Sadik, Ahmed R., et al.
Publicado: (2025)
por: Sadik, Ahmed R., et al.
Publicado: (2025)
Trust-MARL: Trust-Based Multi-Agent Reinforcement Learning Framework for Cooperative On-Ramp Merging Control in Heterogeneous Traffic Flow
por: Pan, Jie, et al.
Publicado: (2025)
por: Pan, Jie, et al.
Publicado: (2025)
Fusion Intelligence: Confluence of Natural and Artificial Intelligence for Enhanced Problem-Solving Efficiency
por: Kalavakonda, Rohan Reddy, et al.
Publicado: (2024)
por: Kalavakonda, Rohan Reddy, et al.
Publicado: (2024)
Multi-Agent Risks from Advanced AI
por: Hammond, Lewis, et al.
Publicado: (2025)
por: Hammond, Lewis, et al.
Publicado: (2025)
Scaling Multiagent Systems with Process Rewards
por: Li, Ed, et al.
Publicado: (2026)
por: Li, Ed, et al.
Publicado: (2026)
From Actions to Understanding: Conformal Interpretability of Temporal Concepts in LLM Agents
por: Padhi, Trilok, et al.
Publicado: (2026)
por: Padhi, Trilok, et al.
Publicado: (2026)
A Gossip-Enhanced Communication Substrate for Agentic AI: Toward Decentralized Coordination in Large-Scale Multi-Agent Systems
por: Khan, Nafiul I., et al.
Publicado: (2025)
por: Khan, Nafiul I., et al.
Publicado: (2025)
Quantum Computing and Neuromorphic Computing for Safe, Reliable, and explainable Multi-Agent Reinforcement Learning: Optimal Control in Autonomous Robotics
por: Taghavi, Mazyar, et al.
Publicado: (2024)
por: Taghavi, Mazyar, et al.
Publicado: (2024)
A Multi-AI Agent System for Autonomous Optimization of Agentic AI Solutions via Iterative Refinement and LLM-Driven Feedback Loops
por: Yuksel, Kamer Ali, et al.
Publicado: (2024)
por: Yuksel, Kamer Ali, et al.
Publicado: (2024)
Visual Reasoning and Multi-Agent Approach in Multimodal Large Language Models (MLLMs): Solving TSP and mTSP Combinatorial Challenges
por: Elhenawy, Mohammed, et al.
Publicado: (2024)
por: Elhenawy, Mohammed, et al.
Publicado: (2024)
Incorporating Large Language Models into Production Systems for Enhanced Task Automation and Flexibility
por: Xia, Yuchen, et al.
Publicado: (2024)
por: Xia, Yuchen, et al.
Publicado: (2024)
Beyond the model: Key differentiators in large language models and multi-agent services
por: Goyal, Muskaan, et al.
Publicado: (2025)
por: Goyal, Muskaan, et al.
Publicado: (2025)
LLM-Ehnanced Holonic Architecture for Ad-Hoc Scalable SoS
por: Ashfaq, Muhammad, et al.
Publicado: (2025)
por: Ashfaq, Muhammad, et al.
Publicado: (2025)
Agentic Witnessing: Pragmatic and Scalable TEE-Enabled Privacy-Preserving Auditing
por: Rowstron, Antony
Publicado: (2026)
por: Rowstron, Antony
Publicado: (2026)
Altruistic Ride Sharing: A Framework for Fair and Sustainable Urban Mobility via Peer-to-Peer Incentives
por: Singh, Divyanshu, et al.
Publicado: (2025)
por: Singh, Divyanshu, et al.
Publicado: (2025)
State and Memory is All You Need for Robust and Reliable AI Agents
por: Muhoberac, Matthew, et al.
Publicado: (2025)
por: Muhoberac, Matthew, et al.
Publicado: (2025)
Adaptive Traffic-Following Scheme for Orderly Distributed Control of Multi-Vehicle Systems
por: Jain, Anahita, et al.
Publicado: (2025)
por: Jain, Anahita, et al.
Publicado: (2025)
The Algorithmic State Architecture (ASA): An Integrated Framework for AI-Enabled Government
por: Engin, Zeynep, et al.
Publicado: (2025)
por: Engin, Zeynep, et al.
Publicado: (2025)
Exploring Agentic Artificial Intelligence Systems: Towards a Typological Framework
por: Wissuchek, Christopher, et al.
Publicado: (2025)
por: Wissuchek, Christopher, et al.
Publicado: (2025)
Computing Threshold Circuits with Bimolecular Void Reactions in Step Chemical Reaction Networks
por: Anderson, Rachel, et al.
Publicado: (2024)
por: Anderson, Rachel, et al.
Publicado: (2024)
SPEAR: An Engineering Case Study of Multi-Agent Coordination for Smart Contract Auditing
por: Chebolu, Indraveni, et al.
Publicado: (2026)
por: Chebolu, Indraveni, et al.
Publicado: (2026)
Ejemplares similares
-
BMG-Q: Localized Bipartite Match Graph Attention Q-Learning for Ride-Pooling Order Dispatch
por: Hu, Yulong, et al.
Publicado: (2025) -
Triple-BERT: Do We Really Need MARL for Order Dispatch on Ride-Sharing Platforms?
por: Zhao, Zijian, et al.
Publicado: (2025) -
LCGuard: Latent Communication Guard for Safe KV Sharing in Multi-Agent Systems
por: Asif, Sadia, et al.
Publicado: (2026) -
Autonomous Traffic Signal Optimization Using Digital Twin and Agentic AI for Real-Time Decision-Making
por: Jan, Salman, et al.
Publicado: (2026) -
ClinEnv: An Interactive Multi-Stage Long Horizon EHR Environment for Agents
por: Lu, Yuxing, et al.
Publicado: (2026)