SPLASH! Sample-efficient Preference-based inverse reinforcement learning for Long-horizon Adversarial tasks from Suboptimal Hierarchical demonstrations
Fuente:
arXiv
Guardado en:
| Autores principales: | Crowley, Peter, Serlin, Zachary, Paine, Tyler, Mann, Makai, Benjamin, Michael, Belta, Calin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Evaluating Collaborative Autonomy in Opposed Environments using Maritime Capture-the-Flag Competitions
por: Beason, Jordan, et al.
Publicado: (2024)
por: Beason, Jordan, et al.
Publicado: (2024)
RRT$^η$: Sampling-based Motion Planning and Control from STL Specifications using Arithmetic-Geometric Mean Robustness
por: Ahmad, Ahmad, et al.
Publicado: (2026)
por: Ahmad, Ahmad, et al.
Publicado: (2026)
Risk-Constrained Belief-Space Optimization for Safe Control under Latent Uncertainty
por: Enwerem, Clinton, et al.
Publicado: (2026)
por: Enwerem, Clinton, et al.
Publicado: (2026)
Ternary Logic Encodings of Temporal Behavior Trees with Application to Control Synthesis
por: Matheu, Ryan, et al.
Publicado: (2026)
por: Matheu, Ryan, et al.
Publicado: (2026)
Learning Safety for Obstacle Avoidance via Control Barrier Functions
por: Liu, Shuo, et al.
Publicado: (2025)
por: Liu, Shuo, et al.
Publicado: (2025)
Iterative Convex Optimization with Control Barrier Functions for Obstacle Avoidance among Polytopes
por: Liu, Shuo, et al.
Publicado: (2026)
por: Liu, Shuo, et al.
Publicado: (2026)
Uncertainty Quantification for Recursive Estimation in Adaptive Safety-Critical Control
por: Cohen, Max H., et al.
Publicado: (2023)
por: Cohen, Max H., et al.
Publicado: (2023)
Long-Horizon Geometry-Aware Navigation among Polytopes via MILP-MPC and Minkowski-Based CBFs
por: Chen, Yi-Hsuan, et al.
Publicado: (2026)
por: Chen, Yi-Hsuan, et al.
Publicado: (2026)
Safety-Critical Planning and Control for Dynamic Obstacle Avoidance Using Control Barrier Functions
por: Liu, Shuo, et al.
Publicado: (2024)
por: Liu, Shuo, et al.
Publicado: (2024)
Comp-LTL: Temporal Logic Planning via Zero-Shot Policy Composition
por: Bergeron, Taylor, et al.
Publicado: (2024)
por: Bergeron, Taylor, et al.
Publicado: (2024)
Accelerated Learning with Linear Temporal Logic using Differentiable Simulation
por: Bozkurt, Alper Kamil, et al.
Publicado: (2025)
por: Bozkurt, Alper Kamil, et al.
Publicado: (2025)
Accelerating Proximal Policy Optimization Learning Using Task Prediction for Solving Environments with Delayed Rewards
por: Ahmad, Ahmad, et al.
Publicado: (2024)
por: Ahmad, Ahmad, et al.
Publicado: (2024)
Auxiliary-Variable Adaptive Control Barrier Functions for Safety Critical Systems
por: Liu, Shuo, et al.
Publicado: (2023)
por: Liu, Shuo, et al.
Publicado: (2023)
Variational Neural Belief Parameterizations for Robust Dexterous Grasping under Multimodal Uncertainty
por: Enwerem, Clinton, et al.
Publicado: (2026)
por: Enwerem, Clinton, et al.
Publicado: (2026)
Safety-Aware Reinforcement Learning for Control via Risk-Sensitive Action-Value Iteration and Quantile Regression
por: Enwerem, Clinton, et al.
Publicado: (2025)
por: Enwerem, Clinton, et al.
Publicado: (2025)
A Model for Multi-Agent Autonomy That Uses Opinion Dynamics and Multi-Objective Behavior Optimization
por: Paine, Tyler M., et al.
Publicado: (2023)
por: Paine, Tyler M., et al.
Publicado: (2023)
Control Barrier Functions via Minkowski Operations for Safe Navigation among Polytopic Sets
por: Chen, Yi-Hsuan, et al.
Publicado: (2025)
por: Chen, Yi-Hsuan, et al.
Publicado: (2025)
Learning from Imperfect Demonstrations via Temporal Behavior Tree-Guided Trajectory Repair
por: Puranic, Aniruddh G., et al.
Publicado: (2026)
por: Puranic, Aniruddh G., et al.
Publicado: (2026)
Quantum deep reinforcement learning for humanoid robot navigation task
por: Lokossou, Romerik, et al.
Publicado: (2025)
por: Lokossou, Romerik, et al.
Publicado: (2025)
HDFlow: Hierarchical Diffusion-Flow Planning for Long-horizon Tasks
por: Gireesh, Nandiraju, et al.
Publicado: (2026)
por: Gireesh, Nandiraju, et al.
Publicado: (2026)
Correspondence learning between morphologically different robots via task demonstrations
por: Aktas, Hakan, et al.
Publicado: (2023)
por: Aktas, Hakan, et al.
Publicado: (2023)
Census-Based Population Autonomy For Distributed Robotic Teaming
por: Paine, Tyler M., et al.
Publicado: (2025)
por: Paine, Tyler M., et al.
Publicado: (2025)
Adaptive bias for dissensus in nonlinear opinion dynamics with application to evolutionary division of labor games
por: Paine, Tyler M., et al.
Publicado: (2024)
por: Paine, Tyler M., et al.
Publicado: (2024)
VICtoR: Learning Hierarchical Vision-Instruction Correlation Rewards for Long-horizon Manipulation
por: Hung, Kuo-Han, et al.
Publicado: (2024)
por: Hung, Kuo-Han, et al.
Publicado: (2024)
Autonomous navigation of catheters and guidewires in mechanical thrombectomy using inverse reinforcement learning
por: Robertshaw, Harry, et al.
Publicado: (2024)
por: Robertshaw, Harry, et al.
Publicado: (2024)
From Abstraction to Reality: DARPA's Vision for Robust Sim-to-Real Autonomy
por: Noorani, Erfaun, et al.
Publicado: (2025)
por: Noorani, Erfaun, et al.
Publicado: (2025)
Disentangling perception and reasoning for improving data efficiency in learning cloth manipulation without demonstrations
por: Delehelle, Donatien, et al.
Publicado: (2026)
por: Delehelle, Donatien, et al.
Publicado: (2026)
Model Predictive Control for Magnetically-Actuated Cellbots
por: Kermanshah, Mehdi, et al.
Publicado: (2024)
por: Kermanshah, Mehdi, et al.
Publicado: (2024)
Leveraging LLMs for reward function design in reinforcement learning control tasks
por: Cardenoso, Franklin, et al.
Publicado: (2025)
por: Cardenoso, Franklin, et al.
Publicado: (2025)
Bootstrapping Imitation Learning for Long-horizon Manipulation via Hierarchical Data Collection Space
por: Yang, Jinrong, et al.
Publicado: (2025)
por: Yang, Jinrong, et al.
Publicado: (2025)
PRISM: Complete Online Decentralized Multi-Agent Pathfinding with Rapid Information Sharing using Motion Constraints
por: Lee, Hannah, et al.
Publicado: (2025)
por: Lee, Hannah, et al.
Publicado: (2025)
Rapidly Converging Time-Discounted Ergodicity on Graphs for Active Inspection of Confined Spaces
por: Wong, Benjamin, et al.
Publicado: (2025)
por: Wong, Benjamin, et al.
Publicado: (2025)
Distributed Control Barrier Functions for Safe Multi-Vehicle Navigation in Heterogeneous USV Fleets
por: Paine, Tyler, et al.
Publicado: (2026)
por: Paine, Tyler, et al.
Publicado: (2026)
STEP Planner: Constructing cross-hierarchical subgoal tree as an embodied long-horizon task planner
por: Zhou, Tianxing, et al.
Publicado: (2025)
por: Zhou, Tianxing, et al.
Publicado: (2025)
Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL
por: Zhang, Songyuan, et al.
Publicado: (2025)
por: Zhang, Songyuan, et al.
Publicado: (2025)
Offline robot programming assisted by task demonstration: an AutomationML interoperable solution for glass adhesive application and welding
por: Babcinschi, M., et al.
Publicado: (2024)
por: Babcinschi, M., et al.
Publicado: (2024)
Real-world Reinforcement Learning from Suboptimal Interventions
por: Zhao, Yinuo, et al.
Publicado: (2025)
por: Zhao, Yinuo, et al.
Publicado: (2025)
The intrinsic motivation of reinforcement and imitation learning for sequential tasks
por: Nguyen, Sao Mai
Publicado: (2024)
por: Nguyen, Sao Mai
Publicado: (2024)
ExoStart: Efficient learning for dexterous manipulation with sensorized exoskeleton demonstrations
por: Si, Zilin, et al.
Publicado: (2025)
por: Si, Zilin, et al.
Publicado: (2025)
Diffusion Trajectory-guided Policy for Long-horizon Robot Manipulation
por: Fan, Shichao, et al.
Publicado: (2025)
por: Fan, Shichao, et al.
Publicado: (2025)
Ejemplares similares
-
Evaluating Collaborative Autonomy in Opposed Environments using Maritime Capture-the-Flag Competitions
por: Beason, Jordan, et al.
Publicado: (2024) -
RRT$^η$: Sampling-based Motion Planning and Control from STL Specifications using Arithmetic-Geometric Mean Robustness
por: Ahmad, Ahmad, et al.
Publicado: (2026) -
Risk-Constrained Belief-Space Optimization for Safe Control under Latent Uncertainty
por: Enwerem, Clinton, et al.
Publicado: (2026) -
Ternary Logic Encodings of Temporal Behavior Trees with Application to Control Synthesis
por: Matheu, Ryan, et al.
Publicado: (2026) -
Learning Safety for Obstacle Avoidance via Control Barrier Functions
por: Liu, Shuo, et al.
Publicado: (2025)