Using a single actor to output personalized policy for different intersections
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhou, Kailing, Zhang, Chengwei, Zhan, Furui, Liu, Wanting, Li, Yihong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Enhancing Traffic Signal Control through Model-based Reinforcement Learning and Policy Reuse
por: Li, Yihong, et al.
Publicado: (2025)
por: Li, Yihong, et al.
Publicado: (2025)
Vital Trace: Protocol-Constrained Patient-State Reasoning for Longitudinal Clinical Trajectories
por: Qu, Zhan, et al.
Publicado: (2026)
por: Qu, Zhan, et al.
Publicado: (2026)
Using Analytics on Student Created Data to Content Validate Pedagogical Tools
por: Kos, John, et al.
Publicado: (2023)
por: Kos, John, et al.
Publicado: (2023)
AIR: Unifying Individual and Collective Exploration in Cooperative Multi-Agent Reinforcement Learning
por: Zhou, Guangchong, et al.
Publicado: (2024)
por: Zhou, Guangchong, et al.
Publicado: (2024)
Learning Complex Teamwork Tasks Using a Given Sub-task Decomposition
por: Fosong, Elliot, et al.
Publicado: (2023)
por: Fosong, Elliot, et al.
Publicado: (2023)
Revisiting Multi-Agent World Modeling from a Diffusion-Inspired Perspective
por: Zhang, Yang, et al.
Publicado: (2025)
por: Zhang, Yang, et al.
Publicado: (2025)
Learn as Individuals, Evolve as a Team: Multi-agent LLMs Adaptation in Embodied Environments
por: Li, Xinran, et al.
Publicado: (2025)
por: Li, Xinran, et al.
Publicado: (2025)
A Bargaining-based Approach for Feature Trading in Vertical Federated Learning
por: Cui, Yue, et al.
Publicado: (2024)
por: Cui, Yue, et al.
Publicado: (2024)
KERAP: A Knowledge-Enhanced Reasoning Approach for Accurate Zero-shot Diagnosis Prediction Using Multi-agent LLMs
por: Xie, Yuzhang, et al.
Publicado: (2025)
por: Xie, Yuzhang, et al.
Publicado: (2025)
The Composite Task Challenge for Cooperative Multi-Agent Reinforcement Learning
por: Li, Yurui, et al.
Publicado: (2025)
por: Li, Yurui, et al.
Publicado: (2025)
Not All Turns Matter: Credit Assignment for Multi-Turn Jailbreaking
por: He, Zhida, et al.
Publicado: (2026)
por: He, Zhida, et al.
Publicado: (2026)
Distributed Multi-Agent Coordination Using Multi-Modal Foundation Models
por: Mahmud, Saaduddin, et al.
Publicado: (2025)
por: Mahmud, Saaduddin, et al.
Publicado: (2025)
Kaleidoscope: Learnable Masks for Heterogeneous Multi-agent Reinforcement Learning
por: Li, Xinran, et al.
Publicado: (2024)
por: Li, Xinran, et al.
Publicado: (2024)
Dynamic Pricing in High-Speed Railways Using Multi-Agent Reinforcement Learning
por: Villarrubia-Martin, Enrique Adrian, et al.
Publicado: (2025)
por: Villarrubia-Martin, Enrique Adrian, et al.
Publicado: (2025)
On-Time Delivery in Crowdshipping Systems: An Agent-Based Approach Using Streaming Data
por: Dötterl, Jeremias, et al.
Publicado: (2024)
por: Dötterl, Jeremias, et al.
Publicado: (2024)
Interaction Pattern Disentangling for Multi-Agent Reinforcement Learning
por: Liu, Shunyu, et al.
Publicado: (2022)
por: Liu, Shunyu, et al.
Publicado: (2022)
Is Centralized Training with Decentralized Execution Framework Centralized Enough for MARL?
por: Zhou, Yihe, et al.
Publicado: (2023)
por: Zhou, Yihe, et al.
Publicado: (2023)
ProAgent: Building Proactive Cooperative Agents with Large Language Models
por: Zhang, Ceyao, et al.
Publicado: (2023)
por: Zhang, Ceyao, et al.
Publicado: (2023)
KernelSkill: A Multi-Agent Framework for GPU Kernel Optimization
por: Sun, Qitong, et al.
Publicado: (2026)
por: Sun, Qitong, et al.
Publicado: (2026)
Decentralized Transformers with Centralized Aggregation are Sample-Efficient Multi-Agent World Models
por: Zhang, Yang, et al.
Publicado: (2024)
por: Zhang, Yang, et al.
Publicado: (2024)
QSIM: Mitigating Overestimation in Multi-Agent Reinforcement Learning via Action Similarity Weighted Q-Learning
por: Li, Yuanjun, et al.
Publicado: (2026)
por: Li, Yuanjun, et al.
Publicado: (2026)
Exponential Topology-enabled Scalable Communication in Multi-agent Reinforcement Learning
por: Li, Xinran, et al.
Publicado: (2025)
por: Li, Xinran, et al.
Publicado: (2025)
Using Deep Q-Learning to Dynamically Toggle between Push/Pull Actions in Computational Trust Mechanisms
por: Lygizou, Zoi, et al.
Publicado: (2024)
por: Lygizou, Zoi, et al.
Publicado: (2024)
The Overcooked Generalisation Challenge: Evaluating Cooperation with Novel Partners in Unknown Environments Using Unsupervised Environment Design
por: Ruhdorfer, Constantin, et al.
Publicado: (2024)
por: Ruhdorfer, Constantin, et al.
Publicado: (2024)
RiskQ: Risk-sensitive Multi-Agent Reinforcement Learning Value Factorization
por: Shen, Siqi, et al.
Publicado: (2023)
por: Shen, Siqi, et al.
Publicado: (2023)
Thought Communication in Multiagent Collaboration
por: Zheng, Yujia, et al.
Publicado: (2025)
por: Zheng, Yujia, et al.
Publicado: (2025)
The Law of Multi-Model Collaboration: Scaling Limits of Model Ensembling for Large Language Models
por: Lu, Dakuan, et al.
Publicado: (2025)
por: Lu, Dakuan, et al.
Publicado: (2025)
Efficient Multi-agent Reinforcement Learning by Planning
por: Liu, Qihan, et al.
Publicado: (2024)
por: Liu, Qihan, et al.
Publicado: (2024)
PMAT: Optimizing Action Generation Order in Multi-Agent Reinforcement Learning
por: Hu, Kun, et al.
Publicado: (2025)
por: Hu, Kun, et al.
Publicado: (2025)
GoAgent: Group-of-Agents Communication Topology Generation for LLM-based Multi-Agent Systems
por: Chen, Hongjiang, et al.
Publicado: (2026)
por: Chen, Hongjiang, et al.
Publicado: (2026)
Planner Matters! An Efficient and Unbalanced Multi-agent Collaboration Framework for Long-horizon Planning
por: Wu, Wenyi, et al.
Publicado: (2026)
por: Wu, Wenyi, et al.
Publicado: (2026)
WideSeek-R1: Exploring Width Scaling for Broad Information Seeking via Multi-Agent Reinforcement Learning
por: Xu, Zelai, et al.
Publicado: (2026)
por: Xu, Zelai, et al.
Publicado: (2026)
Towards Global Optimality in Cooperative MARL with the Transformation And Distillation Framework
por: Ye, Jianing, et al.
Publicado: (2022)
por: Ye, Jianing, et al.
Publicado: (2022)
Context Learning for Multi-Agent Discussion
por: Hua, Xingyuan, et al.
Publicado: (2026)
por: Hua, Xingyuan, et al.
Publicado: (2026)
Flow: Modularized Agentic Workflow Automation
por: Niu, Boye, et al.
Publicado: (2025)
por: Niu, Boye, et al.
Publicado: (2025)
Single-agent or Multi-agent Systems? Why Not Both?
por: Gao, Mingyan, et al.
Publicado: (2025)
por: Gao, Mingyan, et al.
Publicado: (2025)
How to Train a Leader: Hierarchical Reasoning in Multi-Agent LLMs
por: Estornell, Andrew, et al.
Publicado: (2025)
por: Estornell, Andrew, et al.
Publicado: (2025)
Model-based Multi-agent Reinforcement Learning: Recent Progress and Prospects
por: Wang, Xihuai, et al.
Publicado: (2022)
por: Wang, Xihuai, et al.
Publicado: (2022)
RiskAgent: Synergizing Language Models with Validated Tools for Evidence-Based Risk Prediction
por: Liu, Fenglin, et al.
Publicado: (2025)
por: Liu, Fenglin, et al.
Publicado: (2025)
Empirical Study on Robustness and Resilience in Cooperative Multi-Agent Reinforcement Learning
por: Li, Simin, et al.
Publicado: (2025)
por: Li, Simin, et al.
Publicado: (2025)
Ejemplares similares
-
Enhancing Traffic Signal Control through Model-based Reinforcement Learning and Policy Reuse
por: Li, Yihong, et al.
Publicado: (2025) -
Vital Trace: Protocol-Constrained Patient-State Reasoning for Longitudinal Clinical Trajectories
por: Qu, Zhan, et al.
Publicado: (2026) -
Using Analytics on Student Created Data to Content Validate Pedagogical Tools
por: Kos, John, et al.
Publicado: (2023) -
AIR: Unifying Individual and Collective Exploration in Cooperative Multi-Agent Reinforcement Learning
por: Zhou, Guangchong, et al.
Publicado: (2024) -
Learning Complex Teamwork Tasks Using a Given Sub-task Decomposition
por: Fosong, Elliot, et al.
Publicado: (2023)