Rule-based High-Level Coaching for Goal-Conditioned Reinforcement Learning in Search-and-Rescue UAV Missions Under Limited-Simulation Training
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ramezani, Mahya, Voos, Holger |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Human-Centric Aware UAV Trajectory Planning in Search and Rescue Missions Employing Multi-Objective Reinforcement Learning with AHP and Similarity-Based Experience Replay
von: Ramezani, Mahya, et al.
Veröffentlicht: (2024)
von: Ramezani, Mahya, et al.
Veröffentlicht: (2024)
Motion Control in Multi-Rotor Aerial Robots Using Deep Reinforcement Learning
von: Shetty, Gaurav, et al.
Veröffentlicht: (2025)
von: Shetty, Gaurav, et al.
Veröffentlicht: (2025)
HumanDiffusion: A Vision-Based Diffusion Trajectory Planner with Human-Conditioned Goals for Search and Rescue UAV
von: Batool, Faryal, et al.
Veröffentlicht: (2026)
von: Batool, Faryal, et al.
Veröffentlicht: (2026)
MPC-based Deep Reinforcement Learning Method for Space Robotic Control with Fuel Sloshing Mitigation
von: Ramezani, Mahya, et al.
Veröffentlicht: (2025)
von: Ramezani, Mahya, et al.
Veröffentlicht: (2025)
Goal-Conditioned Data Augmentation for Offline Reinforcement Learning
von: Huang, Xingshuai, et al.
Veröffentlicht: (2024)
von: Huang, Xingshuai, et al.
Veröffentlicht: (2024)
SMORE: Score Models for Offline Goal-Conditioned Reinforcement Learning
von: Sikchi, Harshit, et al.
Veröffentlicht: (2023)
von: Sikchi, Harshit, et al.
Veröffentlicht: (2023)
Exploring the Edges of Latent State Clusters for Goal-Conditioned Reinforcement Learning
von: Duan, Yuanlin, et al.
Veröffentlicht: (2024)
von: Duan, Yuanlin, et al.
Veröffentlicht: (2024)
Task Assignment and Exploration Optimization for Low Altitude UAV Rescue via Generative AI Enhanced Multi-agent Reinforcement Learning
von: Tang, Xin, et al.
Veröffentlicht: (2025)
von: Tang, Xin, et al.
Veröffentlicht: (2025)
Data-Efficient Hierarchical Goal-Conditioned Reinforcement Learning via Normalizing Flows
von: Garg, Shaswat, et al.
Veröffentlicht: (2026)
von: Garg, Shaswat, et al.
Veröffentlicht: (2026)
Research on an Autonomous UAV Search and Rescue System Based on the Improved
von: Chen, Haobin, et al.
Veröffentlicht: (2024)
von: Chen, Haobin, et al.
Veröffentlicht: (2024)
Shrinking POMCP: A Framework for Real-Time UAV Search and Rescue
von: Zhang, Yunuo, et al.
Veröffentlicht: (2024)
von: Zhang, Yunuo, et al.
Veröffentlicht: (2024)
PPO-based Dynamic Control of Uncertain Floating Platforms in the Zero-G Environment
von: Ramezani, Mahya, et al.
Veröffentlicht: (2024)
von: Ramezani, Mahya, et al.
Veröffentlicht: (2024)
Goal-Conditioned Decision Transformer for Multi-Goal Offline Reinforcement Learning
von: Gajewski, Paweł, et al.
Veröffentlicht: (2024)
von: Gajewski, Paweł, et al.
Veröffentlicht: (2024)
Offline Goal-Conditioned Reinforcement Learning for Safety-Critical Tasks with Recovery Policy
von: Cao, Chenyang, et al.
Veröffentlicht: (2024)
von: Cao, Chenyang, et al.
Veröffentlicht: (2024)
General Dynamic Goal Recognition using Goal-Conditioned and Meta Reinforcement Learning
von: Elhadad, Osher, et al.
Veröffentlicht: (2025)
von: Elhadad, Osher, et al.
Veröffentlicht: (2025)
Mission-driven Exploration for Accelerated Deep Reinforcement Learning with Temporal Logic Task Specifications
von: Wang, Jun, et al.
Veröffentlicht: (2023)
von: Wang, Jun, et al.
Veröffentlicht: (2023)
Offline Reinforcement Learning of High-Quality Behaviors Under Robust Style Alignment
von: Petitbois, Mathieu, et al.
Veröffentlicht: (2026)
von: Petitbois, Mathieu, et al.
Veröffentlicht: (2026)
Deep Reinforcement Learning based Autonomous Decision-Making for Cooperative UAVs: A Search and Rescue Real World Application
von: Hickling, Thomas, et al.
Veröffentlicht: (2025)
von: Hickling, Thomas, et al.
Veröffentlicht: (2025)
Revisiting Space Mission Planning: A Reinforcement Learning-Guided Approach for Multi-Debris Rendezvous
von: Bandyopadhyay, Agni, et al.
Veröffentlicht: (2024)
von: Bandyopadhyay, Agni, et al.
Veröffentlicht: (2024)
PDSR: Efficient UAV Deployment for Swift and Accurate Post-Disaster Search and Rescue
von: Abdellatif, Alaa Awad, et al.
Veröffentlicht: (2024)
von: Abdellatif, Alaa Awad, et al.
Veröffentlicht: (2024)
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
von: Park, Seohong, et al.
Veröffentlicht: (2023)
von: Park, Seohong, et al.
Veröffentlicht: (2023)
Training and Simulation of Quadrupedal Robot in Adaptive Stair Climbing for Indoor Firefighting: An End-to-End Reinforcement Learning Approach
von: Huang, Baixiao, et al.
Veröffentlicht: (2026)
von: Huang, Baixiao, et al.
Veröffentlicht: (2026)
STEVE-Audio: Expanding the Goal Conditioning Modalities of Embodied Agents in Minecraft
von: Lenzen, Nicholas, et al.
Veröffentlicht: (2024)
von: Lenzen, Nicholas, et al.
Veröffentlicht: (2024)
Improving Offline Reinforcement Learning with Inaccurate Simulators
von: Hou, Yiwen, et al.
Veröffentlicht: (2024)
von: Hou, Yiwen, et al.
Veröffentlicht: (2024)
A Goal-Oriented Reinforcement Learning-Based Path Planning Algorithm for Modular Self-Reconfigurable Satellites
von: Liu, Bofei, et al.
Veröffentlicht: (2025)
von: Liu, Bofei, et al.
Veröffentlicht: (2025)
Slot-MPC: Goal-Conditioned Model Predictive Control with Object-Centric Representations
von: Spieler, Jonathan, et al.
Veröffentlicht: (2026)
von: Spieler, Jonathan, et al.
Veröffentlicht: (2026)
Diffusion Policies with Value-Conditional Optimization for Offline Reinforcement Learning
von: Ma, Yunchang, et al.
Veröffentlicht: (2025)
von: Ma, Yunchang, et al.
Veröffentlicht: (2025)
Safe Reinforcement Learning in a Simulated Robotic Arm
von: Kovač, Luka, et al.
Veröffentlicht: (2023)
von: Kovač, Luka, et al.
Veröffentlicht: (2023)
Online Training and Pruning of Deep Reinforcement Learning Networks
von: Guenter, Valentin Frank Ingmar, et al.
Veröffentlicht: (2025)
von: Guenter, Valentin Frank Ingmar, et al.
Veröffentlicht: (2025)
ACL-QL: Adaptive Conservative Level in Q-Learning for Offline Reinforcement Learning
von: Wu, Kun, et al.
Veröffentlicht: (2024)
von: Wu, Kun, et al.
Veröffentlicht: (2024)
Efficient Virtuoso: A Latent Diffusion Transformer Model for Goal-Conditioned Trajectory Planning
von: Guillen-Perez, Antonio
Veröffentlicht: (2025)
von: Guillen-Perez, Antonio
Veröffentlicht: (2025)
A Safety Modulator Actor-Critic Method in Model-Free Safe Reinforcement Learning and Application in UAV Hovering
von: Qi, Qihan, et al.
Veröffentlicht: (2024)
von: Qi, Qihan, et al.
Veröffentlicht: (2024)
An LLM-based Framework for Human-Swarm Teaming Cognition in Disaster Search and Rescue
von: Ji, Kailun, et al.
Veröffentlicht: (2025)
von: Ji, Kailun, et al.
Veröffentlicht: (2025)
Rating-based Reinforcement Learning
von: White, Devin, et al.
Veröffentlicht: (2023)
von: White, Devin, et al.
Veröffentlicht: (2023)
Learning World Models for Unconstrained Goal Navigation
von: Duan, Yuanlin, et al.
Veröffentlicht: (2024)
von: Duan, Yuanlin, et al.
Veröffentlicht: (2024)
TreeIRL: Safe Urban Driving with Tree Search and Inverse Reinforcement Learning
von: Tomov, Momchil S., et al.
Veröffentlicht: (2025)
von: Tomov, Momchil S., et al.
Veröffentlicht: (2025)
Hyper-GoalNet: Goal-Conditioned Manipulation Policy Learning with HyperNetworks
von: Zhou, Pei, et al.
Veröffentlicht: (2025)
von: Zhou, Pei, et al.
Veröffentlicht: (2025)
Unsupervised Meta-Testing with Conditional Neural Processes for Hybrid Meta-Reinforcement Learning
von: Ada, Suzan Ece, et al.
Veröffentlicht: (2025)
von: Ada, Suzan Ece, et al.
Veröffentlicht: (2025)
Beyond Hard Constraints: Budget-Conditioned Reachability For Safe Offline Reinforcement Learning
von: Brahmanage, Janaka Chathuranga, et al.
Veröffentlicht: (2026)
von: Brahmanage, Janaka Chathuranga, et al.
Veröffentlicht: (2026)
A Plug-and-Play Fully On-the-Job Real-Time Reinforcement Learning Algorithm for a Direct-Drive Tandem-Wing Experiment Platforms Under Multiple Random Operating Conditions
von: Minghao, Zhang, et al.
Veröffentlicht: (2024)
von: Minghao, Zhang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Human-Centric Aware UAV Trajectory Planning in Search and Rescue Missions Employing Multi-Objective Reinforcement Learning with AHP and Similarity-Based Experience Replay
von: Ramezani, Mahya, et al.
Veröffentlicht: (2024) -
Motion Control in Multi-Rotor Aerial Robots Using Deep Reinforcement Learning
von: Shetty, Gaurav, et al.
Veröffentlicht: (2025) -
HumanDiffusion: A Vision-Based Diffusion Trajectory Planner with Human-Conditioned Goals for Search and Rescue UAV
von: Batool, Faryal, et al.
Veröffentlicht: (2026) -
MPC-based Deep Reinforcement Learning Method for Space Robotic Control with Fuel Sloshing Mitigation
von: Ramezani, Mahya, et al.
Veröffentlicht: (2025) -
Goal-Conditioned Data Augmentation for Offline Reinforcement Learning
von: Huang, Xingshuai, et al.
Veröffentlicht: (2024)