Human-compatible driving partners through data-regularized self-play reinforcement learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Cornelisse, Daphne, Vinitsky, Eugene |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Building reliable sim driving agents by scaling self-play
por: Cornelisse, Daphne, et al.
Publicado: (2025)
por: Cornelisse, Daphne, et al.
Publicado: (2025)
Learning to Drive in New Cities Without Human Demonstrations
por: Wang, Zilin, et al.
Publicado: (2026)
por: Wang, Zilin, et al.
Publicado: (2026)
Mitigating Metropolitan Carbon Emissions with Dynamic Eco-driving at Scale
por: Jayawardana, Vindula, et al.
Publicado: (2024)
por: Jayawardana, Vindula, et al.
Publicado: (2024)
Generalizing Cooperative Eco-driving via Multi-residual Task Learning
por: Jayawardana, Vindula, et al.
Publicado: (2024)
por: Jayawardana, Vindula, et al.
Publicado: (2024)
InteRACT: Transformer Models for Human Intent Prediction Conditioned on Robot Actions
por: Kedia, Kushal, et al.
Publicado: (2023)
por: Kedia, Kushal, et al.
Publicado: (2023)
Scalable Multi-Agent Reinforcement Learning for Warehouse Logistics with Robotic and Human Co-Workers
por: Krnjaic, Aleksandar, et al.
Publicado: (2022)
por: Krnjaic, Aleksandar, et al.
Publicado: (2022)
Superhuman Safe and Agile Racing through Multi-Agent Reinforcement Learning
por: Geles, Ismail, et al.
Publicado: (2026)
por: Geles, Ismail, et al.
Publicado: (2026)
Enhancing Aerial Combat Tactics through Hierarchical Multi-Agent Reinforcement Learning
por: Selmonaj, Ardian, et al.
Publicado: (2025)
por: Selmonaj, Ardian, et al.
Publicado: (2025)
Human Implicit Preference-Based Policy Fine-tuning for Multi-Agent Reinforcement Learning in USV Swarm
por: Kim, Hyeonjun, et al.
Publicado: (2025)
por: Kim, Hyeonjun, et al.
Publicado: (2025)
Optimizing Crowd-Aware Multi-Agent Path Finding through Local Communication with Graph Neural Networks
por: Pham, Phu, et al.
Publicado: (2023)
por: Pham, Phu, et al.
Publicado: (2023)
Video Game Level Design as a Multi-Agent Reinforcement Learning Problem
por: Earle, Sam, et al.
Publicado: (2025)
por: Earle, Sam, et al.
Publicado: (2025)
Artificial Intelligence for Modeling and Simulation of Mixed Automated and Human Traffic
por: Rahmani, Saeed, et al.
Publicado: (2026)
por: Rahmani, Saeed, et al.
Publicado: (2026)
Learning to Cooperate with Humans using Generative Agents
por: Liang, Yancheng, et al.
Publicado: (2024)
por: Liang, Yancheng, et al.
Publicado: (2024)
HiMAP: Learning Heuristics-Informed Policies for Large-Scale Multi-Agent Pathfinding
por: Tang, Huijie, et al.
Publicado: (2024)
por: Tang, Huijie, et al.
Publicado: (2024)
Ensembling Prioritized Hybrid Policies for Multi-agent Pathfinding
por: Tang, Huijie, et al.
Publicado: (2024)
por: Tang, Huijie, et al.
Publicado: (2024)
Distributed Autonomous Swarm Formation for Dynamic Network Bridging
por: Galliera, Raffaele, et al.
Publicado: (2024)
por: Galliera, Raffaele, et al.
Publicado: (2024)
Deploying Ten Thousand Robots: Scalable Imitation Learning for Lifelong Multi-Agent Path Finding
por: Jiang, He, et al.
Publicado: (2024)
por: Jiang, He, et al.
Publicado: (2024)
Assigning Credit with Partial Reward Decoupling in Multi-Agent Proximal Policy Optimization
por: Kapoor, Aditya, et al.
Publicado: (2024)
por: Kapoor, Aditya, et al.
Publicado: (2024)
MFC-EQ: Mean-Field Control with Envelope Q-Learning for Moving Decentralized Agents in Formation
por: Lin, Qiushi, et al.
Publicado: (2024)
por: Lin, Qiushi, et al.
Publicado: (2024)
Hypergraph-based Motion Generation with Multi-modal Interaction Relational Reasoning
por: Wu, Keshu, et al.
Publicado: (2024)
por: Wu, Keshu, et al.
Publicado: (2024)
TorchDriveEnv: A Reinforcement Learning Benchmark for Autonomous Driving with Reactive, Realistic, and Diverse Non-Playable Characters
por: Lavington, Jonathan Wilder, et al.
Publicado: (2024)
por: Lavington, Jonathan Wilder, et al.
Publicado: (2024)
Signaling and Social Learning in Swarms of Robots
por: Cazenille, Leo, et al.
Publicado: (2024)
por: Cazenille, Leo, et al.
Publicado: (2024)
BonnBot-I Plus: A Bio-diversity Aware Precise Weed Management Robotic Platform
por: Ahmadi, Alireza, et al.
Publicado: (2024)
por: Ahmadi, Alireza, et al.
Publicado: (2024)
Innate-Values-driven Reinforcement Learning based Cooperative Multi-Agent Cognitive Modeling
por: Yang, Qin
Publicado: (2024)
por: Yang, Qin
Publicado: (2024)
Learning Multi-Agent Loco-Manipulation for Long-Horizon Quadrupedal Pushing
por: Feng, Yuming, et al.
Publicado: (2024)
por: Feng, Yuming, et al.
Publicado: (2024)
Multi-Agent Inverse Reinforcement Learning in Real World Unstructured Pedestrian Crowds
por: Chandra, Rohan, et al.
Publicado: (2024)
por: Chandra, Rohan, et al.
Publicado: (2024)
Controlling Behavioral Diversity in Multi-Agent Reinforcement Learning
por: Bettini, Matteo, et al.
Publicado: (2024)
por: Bettini, Matteo, et al.
Publicado: (2024)
Multi-Agent Deep Q-Network with Layer-based Communication Channel for Autonomous Internal Logistics Vehicle Scheduling in Smart Manufacturing
por: Feizabadi, Mohammad, et al.
Publicado: (2024)
por: Feizabadi, Mohammad, et al.
Publicado: (2024)
Ready, Bid, Go! On-Demand Delivery Using Fleets of Drones with Unknown, Heterogeneous Energy Storage Constraints
por: Talamali, Mohamed S., et al.
Publicado: (2025)
por: Talamali, Mohamed S., et al.
Publicado: (2025)
Communication-Free Collective Navigation for a Swarm of UAVs via LiDAR-Based Deep Reinforcement Learning
por: Choi, Myong-Yol, et al.
Publicado: (2026)
por: Choi, Myong-Yol, et al.
Publicado: (2026)
Compositional Coordination for Multi-Robot Teams with Large Language Models
por: Huang, Zhehui, et al.
Publicado: (2025)
por: Huang, Zhehui, et al.
Publicado: (2025)
Redistributing Rewards Across Time and Agents for Multi-Agent Reinforcement Learning
por: Kapoor, Aditya, et al.
Publicado: (2025)
por: Kapoor, Aditya, et al.
Publicado: (2025)
MARFT: Multi-Agent Reinforcement Fine-Tuning
por: Liao, Junwei, et al.
Publicado: (2025)
por: Liao, Junwei, et al.
Publicado: (2025)
SkyRover: A Modular Simulator for Cross-Domain Pathfinding
por: Ma, Wenhui, et al.
Publicado: (2025)
por: Ma, Wenhui, et al.
Publicado: (2025)
Multi-Agent Inverse Q-Learning from Demonstrations
por: Haynam, Nathaniel, et al.
Publicado: (2025)
por: Haynam, Nathaniel, et al.
Publicado: (2025)
System Neural Diversity: Measuring Behavioral Heterogeneity in Multi-Agent Learning
por: Bettini, Matteo, et al.
Publicado: (2023)
por: Bettini, Matteo, et al.
Publicado: (2023)
Reliable and Efficient Multi-Agent Coordination via Graph Neural Network Variational Autoencoders
por: Meng, Yue, et al.
Publicado: (2025)
por: Meng, Yue, et al.
Publicado: (2025)
Towards Learning Scalable Agile Dynamic Motion Planning for Robosoccer Teams with Policy Optimization
por: Ho, Brandon, et al.
Publicado: (2025)
por: Ho, Brandon, et al.
Publicado: (2025)
Low-Rank Agent-Specific Adaptation (LoRASA) for Multi-Agent Policy Learning
por: Zhang, Beining, et al.
Publicado: (2025)
por: Zhang, Beining, et al.
Publicado: (2025)
Deep Reinforcement Learning for Multi-Agent Coordination
por: Aina, Kehinde O., et al.
Publicado: (2025)
por: Aina, Kehinde O., et al.
Publicado: (2025)
Ejemplares similares
-
Building reliable sim driving agents by scaling self-play
por: Cornelisse, Daphne, et al.
Publicado: (2025) -
Learning to Drive in New Cities Without Human Demonstrations
por: Wang, Zilin, et al.
Publicado: (2026) -
Mitigating Metropolitan Carbon Emissions with Dynamic Eco-driving at Scale
por: Jayawardana, Vindula, et al.
Publicado: (2024) -
Generalizing Cooperative Eco-driving via Multi-residual Task Learning
por: Jayawardana, Vindula, et al.
Publicado: (2024) -
InteRACT: Transformer Models for Human Intent Prediction Conditioned on Robot Actions
por: Kedia, Kushal, et al.
Publicado: (2023)