Bi-level Mean Field: Dynamic Grouping for Large-Scale MARL
Fuente:
arXiv
Salvato in:
| Autori principali: | Zheng, Yuxuan, Zhou, Yihe, Xu, Feiyang, Song, Mingli, Liu, Shunyu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Is Centralized Training with Decentralized Execution Framework Centralized Enough for MARL?
di: Zhou, Yihe, et al.
Pubblicazione: (2023)
di: Zhou, Yihe, et al.
Pubblicazione: (2023)
Temporal Prototype-Aware Learning for Active Voltage Control on Power Distribution Networks
di: Xu, Feiyang, et al.
Pubblicazione: (2024)
di: Xu, Feiyang, et al.
Pubblicazione: (2024)
COLA: Cross-city Mobility Transformer for Human Trajectory Simulation
di: Wang, Yu, et al.
Pubblicazione: (2024)
di: Wang, Yu, et al.
Pubblicazione: (2024)
A2PO: Towards Effective Offline Reinforcement Learning from an Advantage-aware Perspective
di: Qing, Yunpeng, et al.
Pubblicazione: (2024)
di: Qing, Yunpeng, et al.
Pubblicazione: (2024)
Interaction Pattern Disentangling for Multi-Agent Reinforcement Learning
di: Liu, Shunyu, et al.
Pubblicazione: (2022)
di: Liu, Shunyu, et al.
Pubblicazione: (2022)
Parallelized Planning-Acting for Efficient LLM-based Multi-Agent Systems in Minecraft
di: Li, Yaoru, et al.
Pubblicazione: (2025)
di: Li, Yaoru, et al.
Pubblicazione: (2025)
Odyssey: Empowering Minecraft Agents with Open-World Skills
di: Liu, Shunyu, et al.
Pubblicazione: (2024)
di: Liu, Shunyu, et al.
Pubblicazione: (2024)
GraphScout: Empowering Large Language Models with Intrinsic Exploration Ability for Agentic Graph Reasoning
di: Ying, Yuchen, et al.
Pubblicazione: (2026)
di: Ying, Yuchen, et al.
Pubblicazione: (2026)
Unveiling Global Interactive Patterns across Graphs: Towards Interpretable Graph Neural Networks
di: Wang, Yuwen, et al.
Pubblicazione: (2024)
di: Wang, Yuwen, et al.
Pubblicazione: (2024)
SeRL: Self-Play Reinforcement Learning for Large Language Models with Limited Data
di: Fang, Wenkai, et al.
Pubblicazione: (2025)
di: Fang, Wenkai, et al.
Pubblicazione: (2025)
A Survey on Explainable Reinforcement Learning: Concepts, Algorithms, Challenges
di: Qing, Yunpeng, et al.
Pubblicazione: (2022)
di: Qing, Yunpeng, et al.
Pubblicazione: (2022)
Ask-AC: An Initiative Advisor-in-the-Loop Actor-Critic Framework
di: Liu, Shunyu, et al.
Pubblicazione: (2022)
di: Liu, Shunyu, et al.
Pubblicazione: (2022)
Reinforced Model Merging
di: Han, Jiaqi, et al.
Pubblicazione: (2025)
di: Han, Jiaqi, et al.
Pubblicazione: (2025)
Simple Graph Condensation
di: Xiao, Zhenbang, et al.
Pubblicazione: (2024)
di: Xiao, Zhenbang, et al.
Pubblicazione: (2024)
Spatiotemporal-Augmented Graph Neural Networks for Human Mobility Simulation
di: Wang, Yu, et al.
Pubblicazione: (2023)
di: Wang, Yu, et al.
Pubblicazione: (2023)
Breaking the Exploration Bottleneck: Rubric-Scaffolded Reinforcement Learning for General LLM Reasoning
di: Zhou, Yang, et al.
Pubblicazione: (2025)
di: Zhou, Yang, et al.
Pubblicazione: (2025)
Transmission Interface Power Flow Adjustment: A Deep Reinforcement Learning Approach based on Multi-task Attribution Map
di: Liu, Shunyu, et al.
Pubblicazione: (2024)
di: Liu, Shunyu, et al.
Pubblicazione: (2024)
Reasoning with Reinforced Functional Token Tuning
di: Zhang, Kongcheng, et al.
Pubblicazione: (2025)
di: Zhang, Kongcheng, et al.
Pubblicazione: (2025)
Overcoming Environmental Meta-Stationarity in MARL via Adaptive Curriculum and Counterfactual Group Advantage
di: Jin, Weiqiang, et al.
Pubblicazione: (2025)
di: Jin, Weiqiang, et al.
Pubblicazione: (2025)
Tree of Preferences for Diversified Recommendation
di: Yuan, Hanyang, et al.
Pubblicazione: (2025)
di: Yuan, Hanyang, et al.
Pubblicazione: (2025)
MARL Warehouse Robots
di: Allman, Price, et al.
Pubblicazione: (2025)
di: Allman, Price, et al.
Pubblicazione: (2025)
Consistent Paths Lead to Truth: Self-Rewarding Reinforcement Learning for LLM Reasoning
di: Zhang, Kongcheng, et al.
Pubblicazione: (2025)
di: Zhang, Kongcheng, et al.
Pubblicazione: (2025)
Curriculum Negative Mining For Temporal Networks
di: Chen, Ziyue, et al.
Pubblicazione: (2024)
di: Chen, Ziyue, et al.
Pubblicazione: (2024)
Training-free Heterogeneous Model Merging
di: Xu, Zhengqi, et al.
Pubblicazione: (2024)
di: Xu, Zhengqi, et al.
Pubblicazione: (2024)
Towards Efficient LLM-aware Heterogeneous Graph Learning
di: Li, Wenda, et al.
Pubblicazione: (2025)
di: Li, Wenda, et al.
Pubblicazione: (2025)
BiLD: Bi-directional Logits Difference Loss for Large Language Model Distillation
di: Li, Minchong, et al.
Pubblicazione: (2024)
di: Li, Minchong, et al.
Pubblicazione: (2024)
Memory-Efficient Gradient Unrolling for Large-Scale Bi-level Optimization
di: Shen, Qianli, et al.
Pubblicazione: (2024)
di: Shen, Qianli, et al.
Pubblicazione: (2024)
Iterative Semantic Reasoning from Individual to Group Interests for Generative Recommendation with LLMs
di: Zhu, Xiaofei, et al.
Pubblicazione: (2026)
di: Zhu, Xiaofei, et al.
Pubblicazione: (2026)
Guidelines for Applying RL and MARL in Cybersecurity Applications
di: Mavroudis, Vasilios, et al.
Pubblicazione: (2025)
di: Mavroudis, Vasilios, et al.
Pubblicazione: (2025)
Physics-informed Diffusion Generation for Geomagnetic Map Interpolation
di: Li, Wenda, et al.
Pubblicazione: (2026)
di: Li, Wenda, et al.
Pubblicazione: (2026)
Learning Macroeconomic Policies through Dynamic Stackelberg Mean-Field Games
di: Mi, Qirui, et al.
Pubblicazione: (2024)
di: Mi, Qirui, et al.
Pubblicazione: (2024)
Replay Failures as Successes: Sample-Efficient Reinforcement Learning for Instruction Following
di: Zhang, Kongcheng, et al.
Pubblicazione: (2025)
di: Zhang, Kongcheng, et al.
Pubblicazione: (2025)
MF-LLM: Simulating Population Decision Dynamics via a Mean-Field Large Language Model Framework
di: Mi, Qirui, et al.
Pubblicazione: (2025)
di: Mi, Qirui, et al.
Pubblicazione: (2025)
Topology-Assisted Spatio-Temporal Pattern Disentangling for Scalable MARL in Large-scale Autonomous Traffic Control
di: Li, Rongpeng, et al.
Pubblicazione: (2025)
di: Li, Rongpeng, et al.
Pubblicazione: (2025)
Partially Observable Mean Field Multi-Agent Reinforcement Learning Based on Graph-Attention
di: Yang, Min, et al.
Pubblicazione: (2023)
di: Yang, Min, et al.
Pubblicazione: (2023)
Efficient Large-Scale Traffic Forecasting with Transformers: A Spatial Data Management Perspective
di: Fang, Yuchen, et al.
Pubblicazione: (2024)
di: Fang, Yuchen, et al.
Pubblicazione: (2024)
A Survey on Backdoor Threats in Large Language Models (LLMs): Attacks, Defenses, and Evaluations
di: Zhou, Yihe, et al.
Pubblicazione: (2025)
di: Zhou, Yihe, et al.
Pubblicazione: (2025)
BACE-RUL: A Bi-directional Adversarial Network with Covariate Encoding for Machine Remaining Useful Life Prediction
di: Zhang, Zekai, et al.
Pubblicazione: (2025)
di: Zhang, Zekai, et al.
Pubblicazione: (2025)
Solving a Rubik's Cube Using its Local Graph Structure
di: Yao, Shunyu, et al.
Pubblicazione: (2024)
di: Yao, Shunyu, et al.
Pubblicazione: (2024)
MARL-GPT: Foundation Model for Multi-Agent Reinforcement Learning
di: Nesterova, Maria, et al.
Pubblicazione: (2026)
di: Nesterova, Maria, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Is Centralized Training with Decentralized Execution Framework Centralized Enough for MARL?
di: Zhou, Yihe, et al.
Pubblicazione: (2023) -
Temporal Prototype-Aware Learning for Active Voltage Control on Power Distribution Networks
di: Xu, Feiyang, et al.
Pubblicazione: (2024) -
COLA: Cross-city Mobility Transformer for Human Trajectory Simulation
di: Wang, Yu, et al.
Pubblicazione: (2024) -
A2PO: Towards Effective Offline Reinforcement Learning from an Advantage-aware Perspective
di: Qing, Yunpeng, et al.
Pubblicazione: (2024) -
Interaction Pattern Disentangling for Multi-Agent Reinforcement Learning
di: Liu, Shunyu, et al.
Pubblicazione: (2022)