Factorized Deep Q-Network for Cooperative Multi-Agent Reinforcement Learning in Victim Tagging
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cardei, Maria Ana, Doryab, Afsaneh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Sim-to-reality adaptation for Deep Reinforcement Learning applied to an underwater docking application
von: Chaarani, Alaaeddine, et al.
Veröffentlicht: (2026)
von: Chaarani, Alaaeddine, et al.
Veröffentlicht: (2026)
DM$^2$: Decentralized Multi-Agent Reinforcement Learning for Distribution Matching
von: Wang, Caroline, et al.
Veröffentlicht: (2022)
von: Wang, Caroline, et al.
Veröffentlicht: (2022)
Dynamic UGV-UAV Cooperative Path Planning in Uncertain Environments
von: Nguyen, Ninh, et al.
Veröffentlicht: (2026)
von: Nguyen, Ninh, et al.
Veröffentlicht: (2026)
Automated Generation of MDPs Using Logic Programming and LLMs for Robotic Applications
von: Saccon, Enrico, et al.
Veröffentlicht: (2025)
von: Saccon, Enrico, et al.
Veröffentlicht: (2025)
Modeling Considerations for Developing Deep Space Autonomous Spacecraft and Simulators
von: Agia, Christopher, et al.
Veröffentlicht: (2024)
von: Agia, Christopher, et al.
Veröffentlicht: (2024)
An Automatic Ground Collision Avoidance System with Reinforcement Learning
von: Sevgili, Seyyid Osman, et al.
Veröffentlicht: (2026)
von: Sevgili, Seyyid Osman, et al.
Veröffentlicht: (2026)
A Systematic Study of Multi-Agent Deep Reinforcement Learning for Safe and Robust Autonomous Highway Ramp Entry
von: Schester, Larry, et al.
Veröffentlicht: (2024)
von: Schester, Larry, et al.
Veröffentlicht: (2024)
Ultra-Reduced-Impact-Encased-Logging (URIEL): propose a new method for selective sustainable logging and post-harvest silvicultural treatment in tropical forest using airborne robotics systems
von: Albiero, Daniel, et al.
Veröffentlicht: (2026)
von: Albiero, Daniel, et al.
Veröffentlicht: (2026)
DeepFleet: Multi-Agent Foundation Models for Mobile Robots
von: Agaskar, Ameya, et al.
Veröffentlicht: (2025)
von: Agaskar, Ameya, et al.
Veröffentlicht: (2025)
From Idea to CAD: A Language Model-Driven Multi-Agent System for Collaborative Design
von: Ocker, Felix, et al.
Veröffentlicht: (2025)
von: Ocker, Felix, et al.
Veröffentlicht: (2025)
EMOS: Embodiment-aware Heterogeneous Multi-robot Operating System with LLM Agents
von: Chen, Junting, et al.
Veröffentlicht: (2024)
von: Chen, Junting, et al.
Veröffentlicht: (2024)
Active Inference for an Intelligent Agent in Autonomous Reconnaissance Missions
von: Schubert, Johan, et al.
Veröffentlicht: (2025)
von: Schubert, Johan, et al.
Veröffentlicht: (2025)
Multi-Agent Reinforcement Learning for Deadlock Handling among Autonomous Mobile Robots
von: Müller, Marcel
Veröffentlicht: (2025)
von: Müller, Marcel
Veröffentlicht: (2025)
Scaling Safe Multi-Agent Control for Signal Temporal Logic Specifications
von: Eappen, Joe, et al.
Veröffentlicht: (2025)
von: Eappen, Joe, et al.
Veröffentlicht: (2025)
Coordinating Task Switching in a Robotics Multi-Agent System Using Behavior Trees
von: Haug, Lucas, et al.
Veröffentlicht: (2026)
von: Haug, Lucas, et al.
Veröffentlicht: (2026)
Tool-RoCo: An Agent-as-Tool Self-organization Large Language Model Benchmark in Multi-robot Cooperation
von: Zhang, Ke, et al.
Veröffentlicht: (2025)
von: Zhang, Ke, et al.
Veröffentlicht: (2025)
Multi-Agent Coordination for a Partially Observable and Dynamic Robot Soccer Environment with Limited Communication
von: Affinita, Daniele, et al.
Veröffentlicht: (2024)
von: Affinita, Daniele, et al.
Veröffentlicht: (2024)
Deployment-Time Reliability of Learned Robot Policies
von: Agia, Christopher
Veröffentlicht: (2026)
von: Agia, Christopher
Veröffentlicht: (2026)
Procedural Game Level Design with Deep Reinforcement Learning
von: Özkan, Miraç Buğra
Veröffentlicht: (2025)
von: Özkan, Miraç Buğra
Veröffentlicht: (2025)
HULK: Large-scale Hierarchical Coordination under Continual and Uncertain Temporal Tasks
von: Luo, Qingyuan, et al.
Veröffentlicht: (2026)
von: Luo, Qingyuan, et al.
Veröffentlicht: (2026)
A Micro-Macro Model of Encounter-Driven Information Diffusion in Robot Swarms
von: Catherman, Davis S., et al.
Veröffentlicht: (2026)
von: Catherman, Davis S., et al.
Veröffentlicht: (2026)
Occlusion-Based Object Transportation Around Obstacles With a Swarm of Miniature Robots
von: Queiroz, Breno Cunha, et al.
Veröffentlicht: (2026)
von: Queiroz, Breno Cunha, et al.
Veröffentlicht: (2026)
Differentiable Model Predictive Safety for Heterogeneous Mobility at Urban Intersections
von: Song, Wenzhe, et al.
Veröffentlicht: (2026)
von: Song, Wenzhe, et al.
Veröffentlicht: (2026)
Enhancing Heterogeneous Multi-Agent Cooperation in Decentralized MARL via GNN-driven Intrinsic Rewards
von: Monon, Jahir Sadik, et al.
Veröffentlicht: (2024)
von: Monon, Jahir Sadik, et al.
Veröffentlicht: (2024)
Autono: A ReAct-Based Highly Robust Autonomous Agent Framework
von: Wu, Zihao
Veröffentlicht: (2025)
von: Wu, Zihao
Veröffentlicht: (2025)
Decentralized Aerial Manipulation of a Cable-Suspended Load using Multi-Agent Reinforcement Learning
von: Zeng, Jack, et al.
Veröffentlicht: (2025)
von: Zeng, Jack, et al.
Veröffentlicht: (2025)
Optimal UGV-UAV Cooperative Partitioning and Inspection of Shortest Paths
von: Nguyen, Ninh, et al.
Veröffentlicht: (2026)
von: Nguyen, Ninh, et al.
Veröffentlicht: (2026)
Exploiting Differential Flatness for Efficient Learning-based Model Predictive Control of Constrained Multi-Input Control Affine Systems
von: Farger, Tobias A., et al.
Veröffentlicht: (2026)
von: Farger, Tobias A., et al.
Veröffentlicht: (2026)
Discovering Antagonists in Networks of Systems: Robot Deployment
von: Wenger, Ingeborg, et al.
Veröffentlicht: (2025)
von: Wenger, Ingeborg, et al.
Veröffentlicht: (2025)
End-to-End Low-Level Neural Control of an Industrial-Grade 6D Magnetic Levitation System
von: Hartmann, Philipp, et al.
Veröffentlicht: (2025)
von: Hartmann, Philipp, et al.
Veröffentlicht: (2025)
PilotBench: A Benchmark for General Aviation Agents with Safety Constraints
von: Wu, Yalun, et al.
Veröffentlicht: (2026)
von: Wu, Yalun, et al.
Veröffentlicht: (2026)
Privacy Preserving Multi Agent Path Finding
von: Lehman, Rotem Lev, et al.
Veröffentlicht: (2026)
von: Lehman, Rotem Lev, et al.
Veröffentlicht: (2026)
Agentic Automation of BT-RADS Scoring: End-to-End Multi-Agent System for Standardized Brain Tumor Follow-up Assessment
von: Jabal, Mohamed Sobhi, et al.
Veröffentlicht: (2026)
von: Jabal, Mohamed Sobhi, et al.
Veröffentlicht: (2026)
Bimanual Robot Manipulation via Multi-Agent In-Context Learning
von: Palma, Alessio, et al.
Veröffentlicht: (2026)
von: Palma, Alessio, et al.
Veröffentlicht: (2026)
Tulip Agent -- Enabling LLM-Based Agents to Solve Tasks Using Large Tool Libraries
von: Ocker, Felix, et al.
Veröffentlicht: (2024)
von: Ocker, Felix, et al.
Veröffentlicht: (2024)
Selective Progress-Aware Querying for Human-in-the-Loop Reinforcement Learning
von: Muraleedharan, Anujith, et al.
Veröffentlicht: (2025)
von: Muraleedharan, Anujith, et al.
Veröffentlicht: (2025)
Centrally Coordinated Multi-Agent Reinforcement Learning for Power Grid Topology Control
von: de Mol, Barbera, et al.
Veröffentlicht: (2025)
von: de Mol, Barbera, et al.
Veröffentlicht: (2025)
Federated Multi-Agent Mapping for Planetary Exploration
von: Szatmari, Tiberiu-Ioan, et al.
Veröffentlicht: (2024)
von: Szatmari, Tiberiu-Ioan, et al.
Veröffentlicht: (2024)
ROTATE: Regret-driven Open-ended Training for Ad Hoc Teamwork
von: Wang, Caroline, et al.
Veröffentlicht: (2025)
von: Wang, Caroline, et al.
Veröffentlicht: (2025)
Deep Probabilistic Traversability with Test-time Adaptation for Uncertainty-aware Planetary Rover Navigation
von: Endo, Masafumi, et al.
Veröffentlicht: (2024)
von: Endo, Masafumi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Sim-to-reality adaptation for Deep Reinforcement Learning applied to an underwater docking application
von: Chaarani, Alaaeddine, et al.
Veröffentlicht: (2026) -
DM$^2$: Decentralized Multi-Agent Reinforcement Learning for Distribution Matching
von: Wang, Caroline, et al.
Veröffentlicht: (2022) -
Dynamic UGV-UAV Cooperative Path Planning in Uncertain Environments
von: Nguyen, Ninh, et al.
Veröffentlicht: (2026) -
Automated Generation of MDPs Using Logic Programming and LLMs for Robotic Applications
von: Saccon, Enrico, et al.
Veröffentlicht: (2025) -
Modeling Considerations for Developing Deep Space Autonomous Spacecraft and Simulators
von: Agia, Christopher, et al.
Veröffentlicht: (2024)