SPLASH! Sample-efficient Preference-based inverse reinforcement learning for Long-horizon Adversarial tasks from Suboptimal Hierarchical demonstrations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Crowley, Peter, Serlin, Zachary, Paine, Tyler, Mann, Makai, Benjamin, Michael, Belta, Calin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Evaluating Collaborative Autonomy in Opposed Environments using Maritime Capture-the-Flag Competitions
von: Beason, Jordan, et al.
Veröffentlicht: (2024)
von: Beason, Jordan, et al.
Veröffentlicht: (2024)
RRT$^η$: Sampling-based Motion Planning and Control from STL Specifications using Arithmetic-Geometric Mean Robustness
von: Ahmad, Ahmad, et al.
Veröffentlicht: (2026)
von: Ahmad, Ahmad, et al.
Veröffentlicht: (2026)
Risk-Constrained Belief-Space Optimization for Safe Control under Latent Uncertainty
von: Enwerem, Clinton, et al.
Veröffentlicht: (2026)
von: Enwerem, Clinton, et al.
Veröffentlicht: (2026)
Ternary Logic Encodings of Temporal Behavior Trees with Application to Control Synthesis
von: Matheu, Ryan, et al.
Veröffentlicht: (2026)
von: Matheu, Ryan, et al.
Veröffentlicht: (2026)
Learning Safety for Obstacle Avoidance via Control Barrier Functions
von: Liu, Shuo, et al.
Veröffentlicht: (2025)
von: Liu, Shuo, et al.
Veröffentlicht: (2025)
Iterative Convex Optimization with Control Barrier Functions for Obstacle Avoidance among Polytopes
von: Liu, Shuo, et al.
Veröffentlicht: (2026)
von: Liu, Shuo, et al.
Veröffentlicht: (2026)
Uncertainty Quantification for Recursive Estimation in Adaptive Safety-Critical Control
von: Cohen, Max H., et al.
Veröffentlicht: (2023)
von: Cohen, Max H., et al.
Veröffentlicht: (2023)
Long-Horizon Geometry-Aware Navigation among Polytopes via MILP-MPC and Minkowski-Based CBFs
von: Chen, Yi-Hsuan, et al.
Veröffentlicht: (2026)
von: Chen, Yi-Hsuan, et al.
Veröffentlicht: (2026)
Safety-Critical Planning and Control for Dynamic Obstacle Avoidance Using Control Barrier Functions
von: Liu, Shuo, et al.
Veröffentlicht: (2024)
von: Liu, Shuo, et al.
Veröffentlicht: (2024)
Comp-LTL: Temporal Logic Planning via Zero-Shot Policy Composition
von: Bergeron, Taylor, et al.
Veröffentlicht: (2024)
von: Bergeron, Taylor, et al.
Veröffentlicht: (2024)
Accelerated Learning with Linear Temporal Logic using Differentiable Simulation
von: Bozkurt, Alper Kamil, et al.
Veröffentlicht: (2025)
von: Bozkurt, Alper Kamil, et al.
Veröffentlicht: (2025)
Accelerating Proximal Policy Optimization Learning Using Task Prediction for Solving Environments with Delayed Rewards
von: Ahmad, Ahmad, et al.
Veröffentlicht: (2024)
von: Ahmad, Ahmad, et al.
Veröffentlicht: (2024)
Auxiliary-Variable Adaptive Control Barrier Functions for Safety Critical Systems
von: Liu, Shuo, et al.
Veröffentlicht: (2023)
von: Liu, Shuo, et al.
Veröffentlicht: (2023)
Variational Neural Belief Parameterizations for Robust Dexterous Grasping under Multimodal Uncertainty
von: Enwerem, Clinton, et al.
Veröffentlicht: (2026)
von: Enwerem, Clinton, et al.
Veröffentlicht: (2026)
Safety-Aware Reinforcement Learning for Control via Risk-Sensitive Action-Value Iteration and Quantile Regression
von: Enwerem, Clinton, et al.
Veröffentlicht: (2025)
von: Enwerem, Clinton, et al.
Veröffentlicht: (2025)
A Model for Multi-Agent Autonomy That Uses Opinion Dynamics and Multi-Objective Behavior Optimization
von: Paine, Tyler M., et al.
Veröffentlicht: (2023)
von: Paine, Tyler M., et al.
Veröffentlicht: (2023)
Control Barrier Functions via Minkowski Operations for Safe Navigation among Polytopic Sets
von: Chen, Yi-Hsuan, et al.
Veröffentlicht: (2025)
von: Chen, Yi-Hsuan, et al.
Veröffentlicht: (2025)
Learning from Imperfect Demonstrations via Temporal Behavior Tree-Guided Trajectory Repair
von: Puranic, Aniruddh G., et al.
Veröffentlicht: (2026)
von: Puranic, Aniruddh G., et al.
Veröffentlicht: (2026)
Quantum deep reinforcement learning for humanoid robot navigation task
von: Lokossou, Romerik, et al.
Veröffentlicht: (2025)
von: Lokossou, Romerik, et al.
Veröffentlicht: (2025)
HDFlow: Hierarchical Diffusion-Flow Planning for Long-horizon Tasks
von: Gireesh, Nandiraju, et al.
Veröffentlicht: (2026)
von: Gireesh, Nandiraju, et al.
Veröffentlicht: (2026)
Correspondence learning between morphologically different robots via task demonstrations
von: Aktas, Hakan, et al.
Veröffentlicht: (2023)
von: Aktas, Hakan, et al.
Veröffentlicht: (2023)
Census-Based Population Autonomy For Distributed Robotic Teaming
von: Paine, Tyler M., et al.
Veröffentlicht: (2025)
von: Paine, Tyler M., et al.
Veröffentlicht: (2025)
Adaptive bias for dissensus in nonlinear opinion dynamics with application to evolutionary division of labor games
von: Paine, Tyler M., et al.
Veröffentlicht: (2024)
von: Paine, Tyler M., et al.
Veröffentlicht: (2024)
VICtoR: Learning Hierarchical Vision-Instruction Correlation Rewards for Long-horizon Manipulation
von: Hung, Kuo-Han, et al.
Veröffentlicht: (2024)
von: Hung, Kuo-Han, et al.
Veröffentlicht: (2024)
Autonomous navigation of catheters and guidewires in mechanical thrombectomy using inverse reinforcement learning
von: Robertshaw, Harry, et al.
Veröffentlicht: (2024)
von: Robertshaw, Harry, et al.
Veröffentlicht: (2024)
From Abstraction to Reality: DARPA's Vision for Robust Sim-to-Real Autonomy
von: Noorani, Erfaun, et al.
Veröffentlicht: (2025)
von: Noorani, Erfaun, et al.
Veröffentlicht: (2025)
Disentangling perception and reasoning for improving data efficiency in learning cloth manipulation without demonstrations
von: Delehelle, Donatien, et al.
Veröffentlicht: (2026)
von: Delehelle, Donatien, et al.
Veröffentlicht: (2026)
Model Predictive Control for Magnetically-Actuated Cellbots
von: Kermanshah, Mehdi, et al.
Veröffentlicht: (2024)
von: Kermanshah, Mehdi, et al.
Veröffentlicht: (2024)
Leveraging LLMs for reward function design in reinforcement learning control tasks
von: Cardenoso, Franklin, et al.
Veröffentlicht: (2025)
von: Cardenoso, Franklin, et al.
Veröffentlicht: (2025)
Bootstrapping Imitation Learning for Long-horizon Manipulation via Hierarchical Data Collection Space
von: Yang, Jinrong, et al.
Veröffentlicht: (2025)
von: Yang, Jinrong, et al.
Veröffentlicht: (2025)
PRISM: Complete Online Decentralized Multi-Agent Pathfinding with Rapid Information Sharing using Motion Constraints
von: Lee, Hannah, et al.
Veröffentlicht: (2025)
von: Lee, Hannah, et al.
Veröffentlicht: (2025)
Rapidly Converging Time-Discounted Ergodicity on Graphs for Active Inspection of Confined Spaces
von: Wong, Benjamin, et al.
Veröffentlicht: (2025)
von: Wong, Benjamin, et al.
Veröffentlicht: (2025)
Distributed Control Barrier Functions for Safe Multi-Vehicle Navigation in Heterogeneous USV Fleets
von: Paine, Tyler, et al.
Veröffentlicht: (2026)
von: Paine, Tyler, et al.
Veröffentlicht: (2026)
STEP Planner: Constructing cross-hierarchical subgoal tree as an embodied long-horizon task planner
von: Zhou, Tianxing, et al.
Veröffentlicht: (2025)
von: Zhou, Tianxing, et al.
Veröffentlicht: (2025)
Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL
von: Zhang, Songyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Songyuan, et al.
Veröffentlicht: (2025)
Offline robot programming assisted by task demonstration: an AutomationML interoperable solution for glass adhesive application and welding
von: Babcinschi, M., et al.
Veröffentlicht: (2024)
von: Babcinschi, M., et al.
Veröffentlicht: (2024)
Real-world Reinforcement Learning from Suboptimal Interventions
von: Zhao, Yinuo, et al.
Veröffentlicht: (2025)
von: Zhao, Yinuo, et al.
Veröffentlicht: (2025)
The intrinsic motivation of reinforcement and imitation learning for sequential tasks
von: Nguyen, Sao Mai
Veröffentlicht: (2024)
von: Nguyen, Sao Mai
Veröffentlicht: (2024)
ExoStart: Efficient learning for dexterous manipulation with sensorized exoskeleton demonstrations
von: Si, Zilin, et al.
Veröffentlicht: (2025)
von: Si, Zilin, et al.
Veröffentlicht: (2025)
Diffusion Trajectory-guided Policy for Long-horizon Robot Manipulation
von: Fan, Shichao, et al.
Veröffentlicht: (2025)
von: Fan, Shichao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Evaluating Collaborative Autonomy in Opposed Environments using Maritime Capture-the-Flag Competitions
von: Beason, Jordan, et al.
Veröffentlicht: (2024) -
RRT$^η$: Sampling-based Motion Planning and Control from STL Specifications using Arithmetic-Geometric Mean Robustness
von: Ahmad, Ahmad, et al.
Veröffentlicht: (2026) -
Risk-Constrained Belief-Space Optimization for Safe Control under Latent Uncertainty
von: Enwerem, Clinton, et al.
Veröffentlicht: (2026) -
Ternary Logic Encodings of Temporal Behavior Trees with Application to Control Synthesis
von: Matheu, Ryan, et al.
Veröffentlicht: (2026) -
Learning Safety for Obstacle Avoidance via Control Barrier Functions
von: Liu, Shuo, et al.
Veröffentlicht: (2025)