FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation
Fuente:
arXiv
Guardado en:
| Autores principales: | Jiang, Wenzheng, Wang, Ji, Zhang, Xiongtao, Bao, Weidong, Tan, Cheston, Fan, Flint Xiaofeng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Adaptive Episode Length Adjustment for Multi-agent Reinforcement Learning
por: Yoo, Byunghyun, et al.
Publicado: (2025)
por: Yoo, Byunghyun, et al.
Publicado: (2025)
FedRLHF: A Convergence-Guaranteed Federated Framework for Privacy-Preserving and Personalized RLHF
por: Fan, Flint Xiaofeng, et al.
Publicado: (2024)
por: Fan, Flint Xiaofeng, et al.
Publicado: (2024)
Preserving Disagreement: Architectural Heterogeneity and Coherence Validation in Multi-Agent Policy Simulation
por: Sela, Ariel
Publicado: (2026)
por: Sela, Ariel
Publicado: (2026)
Analysing Factorizations of Action-Value Networks for Cooperative Multi-Agent Reinforcement Learning
por: Castellini, Jacopo, et al.
Publicado: (2019)
por: Castellini, Jacopo, et al.
Publicado: (2019)
A Generative Model of Conspicuous Consumption and Status Signaling
por: Cross, Logan, et al.
Publicado: (2026)
por: Cross, Logan, et al.
Publicado: (2026)
Social Deliberation vs. Social Contracts in Self-Governing Voluntary Organisations
por: Scott, Matthew, et al.
Publicado: (2024)
por: Scott, Matthew, et al.
Publicado: (2024)
MiCRO for Multilateral Negotiations
por: Aguilera-Luzon, David, et al.
Publicado: (2025)
por: Aguilera-Luzon, David, et al.
Publicado: (2025)
Autonomous System Safety Properties with Multi-Machine Hybrid Event-B
por: Banach, Richard
Publicado: (2024)
por: Banach, Richard
Publicado: (2024)
Designing Intelligent Enterprise Agents: A Capability-Aligned Multi-Agent Architecture
por: deVadoss, John
Publicado: (2026)
por: deVadoss, John
Publicado: (2026)
Multi-Agent Reinforcement Learning for Deadlock Handling among Autonomous Mobile Robots
por: Müller, Marcel
Publicado: (2025)
por: Müller, Marcel
Publicado: (2025)
AgentSpawn: Adaptive Multi-Agent Collaboration Through Dynamic Spawning for Long-Horizon Code Generation
por: Costa, Igor
Publicado: (2026)
por: Costa, Igor
Publicado: (2026)
PRIMA: Operational Patterns for Resilient Multi-Agent Research with Verifiable Identity and Convergent Feedback
por: Annapureddy, Sasank
Publicado: (2026)
por: Annapureddy, Sasank
Publicado: (2026)
Towards General Negotiation Strategies with End-to-End Reinforcement Learning
por: Renting, Bram M., et al.
Publicado: (2024)
por: Renting, Bram M., et al.
Publicado: (2024)
Accelerating Earth Science Discovery via Multi-Agent LLM Systems
por: Pantiukhin, Dmitrii, et al.
Publicado: (2025)
por: Pantiukhin, Dmitrii, et al.
Publicado: (2025)
Difference Rewards Policy Gradients
por: Castellini, Jacopo, et al.
Publicado: (2020)
por: Castellini, Jacopo, et al.
Publicado: (2020)
Communication Methods in Multi-Agent Reinforcement Learning
por: Wittner, Christoph
Publicado: (2026)
por: Wittner, Christoph
Publicado: (2026)
Privacy Preserving Multi Agent Path Finding
por: Lehman, Rotem Lev, et al.
Publicado: (2026)
por: Lehman, Rotem Lev, et al.
Publicado: (2026)
Event-Triggered Adaptive Consensus for Multi-Robot Task Allocation
por: Aznar, Fidel, et al.
Publicado: (2026)
por: Aznar, Fidel, et al.
Publicado: (2026)
Building Large-Scale Drone Defenses from Small-Team Strategies
por: Douglas, Grant, et al.
Publicado: (2026)
por: Douglas, Grant, et al.
Publicado: (2026)
QTypeMix: Enhancing Multi-Agent Cooperative Strategies through Heterogeneous and Homogeneous Value Decomposition
por: Fu, Songchen, et al.
Publicado: (2024)
por: Fu, Songchen, et al.
Publicado: (2024)
Multi-Agent Reinforcement Learning with Selective State-Space Models
por: Daniel, Jemma, et al.
Publicado: (2024)
por: Daniel, Jemma, et al.
Publicado: (2024)
Dual-Gated Epistemic Time-Dilation: Autonomous Compute Modulation in Asynchronous MARL
por: Jankowski, Igor
Publicado: (2026)
por: Jankowski, Igor
Publicado: (2026)
Cooperative Task Execution in Multi-Agent Systems
por: Karishma, et al.
Publicado: (2024)
por: Karishma, et al.
Publicado: (2024)
Optimal Task Assignment and Path Planning using Conflict-Based Search with Precedence and Temporal Constraints
por: Chong, Yu Quan, et al.
Publicado: (2024)
por: Chong, Yu Quan, et al.
Publicado: (2024)
SkillOps: Managing LLM Agent Skill Libraries as Self-Maintaining Software Ecosystems
por: Pu, Hongji, et al.
Publicado: (2026)
por: Pu, Hongji, et al.
Publicado: (2026)
QRAFTI: An Agentic Framework for Empirical Research in Quantitative Finance
por: Lim, Terence, et al.
Publicado: (2026)
por: Lim, Terence, et al.
Publicado: (2026)
Multi-Agent Empowerment and Emergence of Complex Behavior in Groups
por: Shah, Tristan, et al.
Publicado: (2026)
por: Shah, Tristan, et al.
Publicado: (2026)
Semantics and Spatiality of Emergent Communication
por: Zion, Rotem Ben, et al.
Publicado: (2024)
por: Zion, Rotem Ben, et al.
Publicado: (2024)
Differentiable Model Predictive Safety for Heterogeneous Mobility at Urban Intersections
por: Song, Wenzhe, et al.
Publicado: (2026)
por: Song, Wenzhe, et al.
Publicado: (2026)
Graphon Mean-Field Subsampling for Cooperative Heterogeneous Multi-Agent Reinforcement Learning
por: Anand, Emile, et al.
Publicado: (2026)
por: Anand, Emile, et al.
Publicado: (2026)
Federated Multi-Agent Mapping for Planetary Exploration
por: Szatmari, Tiberiu-Ioan, et al.
Publicado: (2024)
por: Szatmari, Tiberiu-Ioan, et al.
Publicado: (2024)
Agent WARPP: Workflow Adherence via Runtime Parallel Personalization
por: Mazzolenis, Maria Emilia, et al.
Publicado: (2025)
por: Mazzolenis, Maria Emilia, et al.
Publicado: (2025)
DeepFleet: Multi-Agent Foundation Models for Mobile Robots
por: Agaskar, Ameya, et al.
Publicado: (2025)
por: Agaskar, Ameya, et al.
Publicado: (2025)
Multi-Agent Coordination for a Partially Observable and Dynamic Robot Soccer Environment with Limited Communication
por: Affinita, Daniele, et al.
Publicado: (2024)
por: Affinita, Daniele, et al.
Publicado: (2024)
HULK: Large-scale Hierarchical Coordination under Continual and Uncertain Temporal Tasks
por: Luo, Qingyuan, et al.
Publicado: (2026)
por: Luo, Qingyuan, et al.
Publicado: (2026)
Scaling Safe Multi-Agent Control for Signal Temporal Logic Specifications
por: Eappen, Joe, et al.
Publicado: (2025)
por: Eappen, Joe, et al.
Publicado: (2025)
Emergent Coordination in Multi-Agent Language Models
por: Riedl, Christoph
Publicado: (2025)
por: Riedl, Christoph
Publicado: (2025)
A Micro-Macro Model of Encounter-Driven Information Diffusion in Robot Swarms
por: Catherman, Davis S., et al.
Publicado: (2026)
por: Catherman, Davis S., et al.
Publicado: (2026)
Coordinating Task Switching in a Robotics Multi-Agent System Using Behavior Trees
por: Haug, Lucas, et al.
Publicado: (2026)
por: Haug, Lucas, et al.
Publicado: (2026)
Occlusion-Based Object Transportation Around Obstacles With a Swarm of Miniature Robots
por: Queiroz, Breno Cunha, et al.
Publicado: (2026)
por: Queiroz, Breno Cunha, et al.
Publicado: (2026)
Ejemplares similares
-
Adaptive Episode Length Adjustment for Multi-agent Reinforcement Learning
por: Yoo, Byunghyun, et al.
Publicado: (2025) -
FedRLHF: A Convergence-Guaranteed Federated Framework for Privacy-Preserving and Personalized RLHF
por: Fan, Flint Xiaofeng, et al.
Publicado: (2024) -
Preserving Disagreement: Architectural Heterogeneity and Coherence Validation in Multi-Agent Policy Simulation
por: Sela, Ariel
Publicado: (2026) -
Analysing Factorizations of Action-Value Networks for Cooperative Multi-Agent Reinforcement Learning
por: Castellini, Jacopo, et al.
Publicado: (2019) -
A Generative Model of Conspicuous Consumption and Status Signaling
por: Cross, Logan, et al.
Publicado: (2026)