Adaptive Teaching in Heterogeneous Agents: Balancing Surprise in Sparse Reward Scenarios
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Clark, Emma, Ryu, Kanghyun, Mehr, Negar |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Integrating Predictive Motion Uncertainties with Distributionally Robust Risk-Aware Control for Safe Robot Navigation in Crowds
von: Ryu, Kanghyun, et al.
Veröffentlicht: (2024)
von: Ryu, Kanghyun, et al.
Veröffentlicht: (2024)
CurricuLLM: Automatic Task Curricula Design for Learning Complex Robot Skills using Large Language Models
von: Ryu, Kanghyun, et al.
Veröffentlicht: (2024)
von: Ryu, Kanghyun, et al.
Veröffentlicht: (2024)
CRAFT: Coaching Reinforcement Learning Autonomously using Foundation Models for Multi-Robot Coordination Tasks
von: Choi, Seoyeon, et al.
Veröffentlicht: (2025)
von: Choi, Seoyeon, et al.
Veröffentlicht: (2025)
Leveraging Large Language Models for Effective and Explainable Multi-Agent Credit Assignment
von: Nagpal, Kartik, et al.
Veröffentlicht: (2025)
von: Nagpal, Kartik, et al.
Veröffentlicht: (2025)
Surprise Potential as a Measure of Interactivity in Driving Scenarios
von: Ding, Wenhao, et al.
Veröffentlicht: (2025)
von: Ding, Wenhao, et al.
Veröffentlicht: (2025)
Distributed NeRF Learning for Collaborative Multi-Robot Perception
von: Zhao, Hongrui, et al.
Veröffentlicht: (2024)
von: Zhao, Hongrui, et al.
Veröffentlicht: (2024)
Risk-Sensitive Orbital Debris Collision Avoidance using Distributionally Robust Chance Constraints
von: Ryu, Kanghyun, et al.
Veröffentlicht: (2024)
von: Ryu, Kanghyun, et al.
Veröffentlicht: (2024)
Multi-Agent Inverse Reinforcement Learning in Real World Unstructured Pedestrian Crowds
von: Chandra, Rohan, et al.
Veröffentlicht: (2024)
von: Chandra, Rohan, et al.
Veröffentlicht: (2024)
DDAT: Diffusion Policies Enforcing Dynamically Admissible Robot Trajectories
von: Bouvier, Jean-Baptiste, et al.
Veröffentlicht: (2025)
von: Bouvier, Jean-Baptiste, et al.
Veröffentlicht: (2025)
Revisiting Sparse Rewards for Goal-Reaching Reinforcement Learning
von: Vasan, Gautham, et al.
Veröffentlicht: (2024)
von: Vasan, Gautham, et al.
Veröffentlicht: (2024)
TopoNav: Topological Navigation for Efficient Exploration in Sparse Reward Environments
von: Hossain, Jumman, et al.
Veröffentlicht: (2024)
von: Hossain, Jumman, et al.
Veröffentlicht: (2024)
Teaching Robots to Handle Nuclear Waste: A Teleoperation-Based Learning Approach<
von: Lee, Joong-Ku, et al.
Veröffentlicht: (2025)
von: Lee, Joong-Ku, et al.
Veröffentlicht: (2025)
Using Surprise Index for Competency Assessment in Autonomous Decision-Making
von: Ratheesh, Akash, et al.
Veröffentlicht: (2023)
von: Ratheesh, Akash, et al.
Veröffentlicht: (2023)
ETGL-DDPG: A Deep Deterministic Policy Gradient Algorithm for Sparse Reward Continuous Control
von: Futuhi, Ehsan, et al.
Veröffentlicht: (2024)
von: Futuhi, Ehsan, et al.
Veröffentlicht: (2024)
Subwords as Skills: Tokenization for Sparse-Reward Reinforcement Learning
von: Yunis, David, et al.
Veröffentlicht: (2023)
von: Yunis, David, et al.
Veröffentlicht: (2023)
DISCOVER: Automated Curricula for Sparse-Reward Reinforcement Learning
von: Diaz-Bone, Leander, et al.
Veröffentlicht: (2025)
von: Diaz-Bone, Leander, et al.
Veröffentlicht: (2025)
Strategic Decision-Making in Multi-Agent Domains: A Weighted Constrained Potential Dynamic Game Approach
von: Bhatt, Maulik, et al.
Veröffentlicht: (2023)
von: Bhatt, Maulik, et al.
Veröffentlicht: (2023)
SimBEV: A Synthetic Multi-Task Multi-Sensor Driving Data Generation Tool and Dataset
von: Mehr, Goodarz, et al.
Veröffentlicht: (2025)
von: Mehr, Goodarz, et al.
Veröffentlicht: (2025)
Optimal Robotic Assembly Sequence Planning: A Sequential Decision-Making Approach
von: Nagpal, Kartik, et al.
Veröffentlicht: (2023)
von: Nagpal, Kartik, et al.
Veröffentlicht: (2023)
Enabling Option Learning in Sparse Rewards with Hindsight Experience Replay
von: Romio, Gabriel, et al.
Veröffentlicht: (2026)
von: Romio, Gabriel, et al.
Veröffentlicht: (2026)
CAFE-AD: Cross-Scenario Adaptive Feature Enhancement for Trajectory Planning in Autonomous Driving
von: Zhang, Junrui, et al.
Veröffentlicht: (2025)
von: Zhang, Junrui, et al.
Veröffentlicht: (2025)
VLM as Strategist: Adaptive Generation of Safety-critical Testing Scenarios via Guided Diffusion
von: Wu, Xinzheng, et al.
Veröffentlicht: (2025)
von: Wu, Xinzheng, et al.
Veröffentlicht: (2025)
UniGen: Unified Modeling of Initial Agent States and Trajectories for Generating Autonomous Driving Scenarios
von: Mahjourian, Reza, et al.
Veröffentlicht: (2024)
von: Mahjourian, Reza, et al.
Veröffentlicht: (2024)
Confidence-Controlled Exploration: Efficient Sparse-Reward Policy Learning for Robot Navigation
von: Patel, Bhrij, et al.
Veröffentlicht: (2023)
von: Patel, Bhrij, et al.
Veröffentlicht: (2023)
Adaptive Querying for Reward Learning from Human Feedback
von: Anand, Yashwanthi, et al.
Veröffentlicht: (2024)
von: Anand, Yashwanthi, et al.
Veröffentlicht: (2024)
Balancing Signal and Variance: Adaptive Offline RL Post-Training for VLA Flow Models
von: Zhang, Hongyin, et al.
Veröffentlicht: (2025)
von: Zhang, Hongyin, et al.
Veröffentlicht: (2025)
Adaptive Control Strategy for Quadruped Robots in Actuator Degradation Scenarios
von: Wu, Xinyuan, et al.
Veröffentlicht: (2023)
von: Wu, Xinyuan, et al.
Veröffentlicht: (2023)
Rewarding DINO: Predicting Dense Rewards with Vision Foundation Models
von: Krack, Pierre, et al.
Veröffentlicht: (2026)
von: Krack, Pierre, et al.
Veröffentlicht: (2026)
Scenario-Based Curriculum Generation for Multi-Agent Autonomous Driving
von: Brunnbauer, Axel, et al.
Veröffentlicht: (2024)
von: Brunnbauer, Axel, et al.
Veröffentlicht: (2024)
The Dark Side of Rich Rewards: Understanding and Mitigating Noise in VLM Rewards
von: Huang, Sukai, et al.
Veröffentlicht: (2024)
von: Huang, Sukai, et al.
Veröffentlicht: (2024)
RobotKeyframing: Learning Locomotion with High-Level Objectives via Mixture of Dense and Sparse Rewards
von: Zargarbashi, Fatemeh, et al.
Veröffentlicht: (2024)
von: Zargarbashi, Fatemeh, et al.
Veröffentlicht: (2024)
RDAR: Reward-Driven Agent Relevance Estimation for Autonomous Driving
von: Bosio, Carlo, et al.
Veröffentlicht: (2025)
von: Bosio, Carlo, et al.
Veröffentlicht: (2025)
Pretrained Vision-Language-Action Models are Surprisingly Resistant to Forgetting in Continual Learning
von: Liu, Huihan, et al.
Veröffentlicht: (2026)
von: Liu, Huihan, et al.
Veröffentlicht: (2026)
Reward Machine Inference for Robotic Manipulation
von: Baert, Mattijs, et al.
Veröffentlicht: (2024)
von: Baert, Mattijs, et al.
Veröffentlicht: (2024)
ParkDiffusion: Heterogeneous Multi-Agent Multi-Modal Trajectory Prediction for Automated Parking using Diffusion Models
von: Wei, Jiarong, et al.
Veröffentlicht: (2025)
von: Wei, Jiarong, et al.
Veröffentlicht: (2025)
Uncertainty-aware Reward Design Process
von: Yang, Yang, et al.
Veröffentlicht: (2025)
von: Yang, Yang, et al.
Veröffentlicht: (2025)
Matching Multiple Experts: On the Exploitability of Multi-Agent Imitation Learning
von: Bergerault, Antoine, et al.
Veröffentlicht: (2026)
von: Bergerault, Antoine, et al.
Veröffentlicht: (2026)
RAMEN: Real-time Asynchronous Multi-agent Neural Implicit Mapping
von: Zhao, Hongrui, et al.
Veröffentlicht: (2025)
von: Zhao, Hongrui, et al.
Veröffentlicht: (2025)
TACO: Temporal Consensus Optimization for Continual Neural Mapping
von: Zhou, Xunlan, et al.
Veröffentlicht: (2026)
von: Zhou, Xunlan, et al.
Veröffentlicht: (2026)
Learning Speed-Adaptive Walking Agent Using Imitation Learning with Physics-Informed Simulation
von: Chiu, Yi-Hung, et al.
Veröffentlicht: (2024)
von: Chiu, Yi-Hung, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Integrating Predictive Motion Uncertainties with Distributionally Robust Risk-Aware Control for Safe Robot Navigation in Crowds
von: Ryu, Kanghyun, et al.
Veröffentlicht: (2024) -
CurricuLLM: Automatic Task Curricula Design for Learning Complex Robot Skills using Large Language Models
von: Ryu, Kanghyun, et al.
Veröffentlicht: (2024) -
CRAFT: Coaching Reinforcement Learning Autonomously using Foundation Models for Multi-Robot Coordination Tasks
von: Choi, Seoyeon, et al.
Veröffentlicht: (2025) -
Leveraging Large Language Models for Effective and Explainable Multi-Agent Credit Assignment
von: Nagpal, Kartik, et al.
Veröffentlicht: (2025) -
Surprise Potential as a Measure of Interactivity in Driving Scenarios
von: Ding, Wenhao, et al.
Veröffentlicht: (2025)