SMAUG: A Sliding Multidimensional Task Window-Based MARL Framework for Adaptive Real-Time Subtask Recognition
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhang, Wenjing, Zhang, Wei |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Smart Traffic Signals: Comparing MARL and Fixed-Time Strategies
por: Mahato, Saahil
Publicado: (2025)
por: Mahato, Saahil
Publicado: (2025)
Towards Global Optimality in Cooperative MARL with the Transformation And Distillation Framework
por: Ye, Jianing, et al.
Publicado: (2022)
por: Ye, Jianing, et al.
Publicado: (2022)
HyperMARL: Adaptive Hypernetworks for Multi-Agent RL
por: Tessera, Kale-ab Abebe, et al.
Publicado: (2024)
por: Tessera, Kale-ab Abebe, et al.
Publicado: (2024)
Beyond Shallow Behavior: Task-Efficient Value-Based Multi-Task Offline MARL via Skill Discovery
por: Wang, Xun, et al.
Publicado: (2025)
por: Wang, Xun, et al.
Publicado: (2025)
Addressing Situated Teaching Needs: A Multi-Agent Framework for Automated Slide Adaptation
por: Liu, Binglin, et al.
Publicado: (2025)
por: Liu, Binglin, et al.
Publicado: (2025)
A MARL-based Approach for Easing MAS Organization Engineering
por: Soulé, Julien, et al.
Publicado: (2025)
por: Soulé, Julien, et al.
Publicado: (2025)
Learning to Lead Themselves: Agentic AI in MAS using MARL
por: Kamthan, Ansh
Publicado: (2025)
por: Kamthan, Ansh
Publicado: (2025)
Complete Cyclic Subtask Graphs for Tool-Using LLM Agents: Flexibility, Cost, and Bottlenecks in Multi-Agent Workflows
por: Gharzeddine, Luay, et al.
Publicado: (2026)
por: Gharzeddine, Luay, et al.
Publicado: (2026)
Is Centralized Training with Decentralized Execution Framework Centralized Enough for MARL?
por: Zhou, Yihe, et al.
Publicado: (2023)
por: Zhou, Yihe, et al.
Publicado: (2023)
CH-MARL: Constrained Hierarchical Multiagent Reinforcement Learning for Sustainable Maritime Logistics
por: Alqithami, Saad
Publicado: (2025)
por: Alqithami, Saad
Publicado: (2025)
TaskGen: A Task-Based, Memory-Infused Agentic Framework using StrictJSON
por: Tan, John Chong Min, et al.
Publicado: (2024)
por: Tan, John Chong Min, et al.
Publicado: (2024)
EcoFair-CH-MARL: Scalable Constrained Hierarchical Multi-Agent RL with Real-Time Emission Budgets and Fairness Guarantees
por: Alqithami, Saad
Publicado: (2026)
por: Alqithami, Saad
Publicado: (2026)
EDU-MATRIX: A Society-Centric Generative Cognitive Digital Twin Architecture for Secondary Education
por: Zhai, Wenjing, et al.
Publicado: (2026)
por: Zhai, Wenjing, et al.
Publicado: (2026)
Coordination Failure in Cooperative Offline MARL
por: Tilbury, Callum Rhys, et al.
Publicado: (2024)
por: Tilbury, Callum Rhys, et al.
Publicado: (2024)
LLM-Mediated Guidance of MARL Systems
por: Siedler, Philipp D., et al.
Publicado: (2025)
por: Siedler, Philipp D., et al.
Publicado: (2025)
HAMLET: A Hierarchical and Adaptive Multi-Agent Framework for Live Embodied Theatrics
por: Jiang, Shufan, et al.
Publicado: (2025)
por: Jiang, Shufan, et al.
Publicado: (2025)
Task-Aware LLM Council with Adaptive Decision Pathways for Decision Support
por: Zhu, Wei, et al.
Publicado: (2026)
por: Zhu, Wei, et al.
Publicado: (2026)
Investigating Relational State Abstraction in Collaborative MARL
por: Utke, Sharlin, et al.
Publicado: (2024)
por: Utke, Sharlin, et al.
Publicado: (2024)
Learning Partial Action Replacement in Offline MARL
por: Jin, Yue, et al.
Publicado: (2026)
por: Jin, Yue, et al.
Publicado: (2026)
BenchMARL: Benchmarking Multi-Agent Reinforcement Learning
por: Bettini, Matteo, et al.
Publicado: (2023)
por: Bettini, Matteo, et al.
Publicado: (2023)
VillagerAgent: A Graph-Based Multi-Agent Framework for Coordinating Complex Task Dependencies in Minecraft
por: Dong, Yubo, et al.
Publicado: (2024)
por: Dong, Yubo, et al.
Publicado: (2024)
Towards Robust Multi-UAV Collaboration: MARL with Noise-Resilient Communication and Attention Mechanisms
por: Zhao, Zilin, et al.
Publicado: (2025)
por: Zhao, Zilin, et al.
Publicado: (2025)
Sable: a Performant, Efficient and Scalable Sequence Model for MARL
por: Mahjoub, Omayma, et al.
Publicado: (2024)
por: Mahjoub, Omayma, et al.
Publicado: (2024)
JaxMARL: Multi-Agent RL Environments and Algorithms in JAX
por: Rutherford, Alexander, et al.
Publicado: (2023)
por: Rutherford, Alexander, et al.
Publicado: (2023)
Partial Action Replacement: Tackling Distribution Shift in Offline MARL
por: Jin, Yue, et al.
Publicado: (2025)
por: Jin, Yue, et al.
Publicado: (2025)
Decoupled Delay Compensation: Enhancing Pre-trained MARL Policies via Learned Dynamics Filtering
por: Mednikov, Maxim, et al.
Publicado: (2026)
por: Mednikov, Maxim, et al.
Publicado: (2026)
Solving Context Window Overflow in AI Agents
por: Labate, Anton Bulle, et al.
Publicado: (2025)
por: Labate, Anton Bulle, et al.
Publicado: (2025)
MEDCO: Medical Education Copilots Based on A Multi-Agent Framework
por: Wei, Hao, et al.
Publicado: (2024)
por: Wei, Hao, et al.
Publicado: (2024)
Real-Time LaCAM for Real-Time MAPF
por: Liang, Runzhe, et al.
Publicado: (2025)
por: Liang, Runzhe, et al.
Publicado: (2025)
BMW Agents -- A Framework For Task Automation Through Multi-Agent Collaboration
por: Crawford, Noel, et al.
Publicado: (2024)
por: Crawford, Noel, et al.
Publicado: (2024)
Achieving Optimal Tissue Repair Through MARL with Reward Shaping and Curriculum Learning
por: Khan, Muhammad Al-Zafar, et al.
Publicado: (2025)
por: Khan, Muhammad Al-Zafar, et al.
Publicado: (2025)
A Role-Based LLM Framework for Structured Information Extraction from Healthy Food Policies
por: Zhang, Congjing, et al.
Publicado: (2026)
por: Zhang, Congjing, et al.
Publicado: (2026)
Bridging MARL to SARL: An Order-Independent Multi-Agent Transformer via Latent Consensus
por: Zhao, Zijian, et al.
Publicado: (2026)
por: Zhao, Zijian, et al.
Publicado: (2026)
Windowed MAPF with Completeness Guarantees
por: Veerapaneni, Rishi, et al.
Publicado: (2024)
por: Veerapaneni, Rishi, et al.
Publicado: (2024)
DERM-3R: A Resource-Efficient Multimodal Agents Framework for Dermatologic Diagnosis and Treatment in Real-World Clinical Settings
por: Chen, Ziwen, et al.
Publicado: (2026)
por: Chen, Ziwen, et al.
Publicado: (2026)
MetaCogAgent: A Metacognitive Multi-Agent LLM Framework with Self-Aware Task Delegation
por: Wang, Chenyu, et al.
Publicado: (2026)
por: Wang, Chenyu, et al.
Publicado: (2026)
H2-MARL: Multi-Agent Reinforcement Learning for Pareto Optimality in Hospital Capacity Strain and Human Mobility during Epidemic
por: Luo, Xueting, et al.
Publicado: (2025)
por: Luo, Xueting, et al.
Publicado: (2025)
Triple-BERT: Do We Really Need MARL for Order Dispatch on Ride-Sharing Platforms?
por: Zhao, Zijian, et al.
Publicado: (2025)
por: Zhao, Zijian, et al.
Publicado: (2025)
Hierarchical LLM-Based Multi-Agent Framework with Prompt Optimization for Multi-Robot Task Planning
por: Kawabe, Tomoya, et al.
Publicado: (2026)
por: Kawabe, Tomoya, et al.
Publicado: (2026)
MA-GTS: A Multi-Agent Framework for Solving Complex Graph Problems in Real-World Applications
por: Yuan, Zike, et al.
Publicado: (2025)
por: Yuan, Zike, et al.
Publicado: (2025)
Ejemplares similares
-
Smart Traffic Signals: Comparing MARL and Fixed-Time Strategies
por: Mahato, Saahil
Publicado: (2025) -
Towards Global Optimality in Cooperative MARL with the Transformation And Distillation Framework
por: Ye, Jianing, et al.
Publicado: (2022) -
HyperMARL: Adaptive Hypernetworks for Multi-Agent RL
por: Tessera, Kale-ab Abebe, et al.
Publicado: (2024) -
Beyond Shallow Behavior: Task-Efficient Value-Based Multi-Task Offline MARL via Skill Discovery
por: Wang, Xun, et al.
Publicado: (2025) -
Addressing Situated Teaching Needs: A Multi-Agent Framework for Automated Slide Adaptation
por: Liu, Binglin, et al.
Publicado: (2025)