When to Act, Ask, or Learn: Uncertainty-Aware Policy Steering
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yuan, Jessie, Wu, Yilin, Bajcsy, Andrea |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
From Foresight to Forethought: VLM-In-the-Loop Policy Steering via Latent Alignment
von: Wu, Yilin, et al.
Veröffentlicht: (2025)
von: Wu, Yilin, et al.
Veröffentlicht: (2025)
Do What You Say: Steering Vision-Language-Action Models via Runtime Reasoning-Action Alignment Verification
von: Wu, Yilin, et al.
Veröffentlicht: (2025)
von: Wu, Yilin, et al.
Veröffentlicht: (2025)
Uncertainty-aware Latent Safety Filters for Avoiding Out-of-Distribution Failures
von: Seo, Junwon, et al.
Veröffentlicht: (2025)
von: Seo, Junwon, et al.
Veröffentlicht: (2025)
Maximizing Alignment with Minimal Feedback: Efficiently Learning Rewards for Visuomotor Robot Policy Alignment
von: Tian, Ran, et al.
Veröffentlicht: (2024)
von: Tian, Ran, et al.
Veröffentlicht: (2024)
Adapting by Analogy: OOD Generalization of Visuomotor Policies via Functional Correspondence
von: Gupta, Pranay, et al.
Veröffentlicht: (2025)
von: Gupta, Pranay, et al.
Veröffentlicht: (2025)
Generalizing Safety Beyond Collision-Avoidance via Latent-Space Reachability Analysis
von: Nakamura, Kensuke, et al.
Veröffentlicht: (2025)
von: Nakamura, Kensuke, et al.
Veröffentlicht: (2025)
Uncertainty-Aware Deployment of Pre-trained Language-Conditioned Imitation Learning Policies
von: Wu, Bo, et al.
Veröffentlicht: (2024)
von: Wu, Bo, et al.
Veröffentlicht: (2024)
Symmetry-Aware Steering of Equivariant Diffusion Policies: Benefits and Limits
von: Park, Minwoo, et al.
Veröffentlicht: (2025)
von: Park, Minwoo, et al.
Veröffentlicht: (2025)
StressDream: Steering Video World Models for Robust Policy Evaluation and Improvement
von: Seo, Junwon, et al.
Veröffentlicht: (2026)
von: Seo, Junwon, et al.
Veröffentlicht: (2026)
Conformalized Teleoperation: Confidently Mapping Human Inputs to High-Dimensional Robot Actions
von: Zhao, Michelle, et al.
Veröffentlicht: (2024)
von: Zhao, Michelle, et al.
Veröffentlicht: (2024)
Learning When to Act: Communication-Efficient Reinforcement Learning via Run-Time Assurance
von: Haroon, Adam, et al.
Veröffentlicht: (2026)
von: Haroon, Adam, et al.
Veröffentlicht: (2026)
Steering Your Diffusion Policy with Latent Space Reinforcement Learning
von: Wagenmaker, Andrew, et al.
Veröffentlicht: (2025)
von: Wagenmaker, Andrew, et al.
Veröffentlicht: (2025)
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation
von: Danesh, Mohamad H., et al.
Veröffentlicht: (2025)
von: Danesh, Mohamad H., et al.
Veröffentlicht: (2025)
AnySafe: Adapting Latent Safety Filters at Runtime via Safety Constraint Parameterization in the Latent Space
von: Agrawal, Sankalp, et al.
Veröffentlicht: (2025)
von: Agrawal, Sankalp, et al.
Veröffentlicht: (2025)
Your Learned Constraint is Secretly a Backward Reachable Tube
von: Qadri, Mohamad, et al.
Veröffentlicht: (2025)
von: Qadri, Mohamad, et al.
Veröffentlicht: (2025)
Latent Policy Steering through One-Step Flow Policies
von: Im, Hokyun, et al.
Veröffentlicht: (2026)
von: Im, Hokyun, et al.
Veröffentlicht: (2026)
Learning When to Ask: Simulation-Trained Humanoids for Mental-Health Diagnosis
von: Cenacchi, Filippo, et al.
Veröffentlicht: (2025)
von: Cenacchi, Filippo, et al.
Veröffentlicht: (2025)
EquivAct: SIM(3)-Equivariant Visuomotor Policies beyond Rigid Object Manipulation
von: Yang, Jingyun, et al.
Veröffentlicht: (2023)
von: Yang, Jingyun, et al.
Veröffentlicht: (2023)
Position: Good Embodied Reward Models Need Bad Behavior Data
von: Tian, Ran, et al.
Veröffentlicht: (2026)
von: Tian, Ran, et al.
Veröffentlicht: (2026)
How to Train Your Latent Control Barrier Function: Smooth Safety Filtering Under Hard-to-Model Constraints
von: Nakamura, Kensuke, et al.
Veröffentlicht: (2025)
von: Nakamura, Kensuke, et al.
Veröffentlicht: (2025)
Ask1: Development and Reinforcement Learning-Based Control of a Custom Quadruped Robot
von: Zhang, Yang, et al.
Veröffentlicht: (2024)
von: Zhang, Yang, et al.
Veröffentlicht: (2024)
Conformal Decision Theory: Safe Autonomous Decisions from Imperfect Predictions
von: Lekeufack, Jordan, et al.
Veröffentlicht: (2023)
von: Lekeufack, Jordan, et al.
Veröffentlicht: (2023)
Act-Observe-Rewrite: Multimodal Coding Agents as In-Context Policy Learners for Robot Manipulation
von: Kumar, Vaishak
Veröffentlicht: (2026)
von: Kumar, Vaishak
Veröffentlicht: (2026)
Conformalized Interactive Imitation Learning: Handling Expert Shift and Intermittent Feedback
von: Zhao, Michelle, et al.
Veröffentlicht: (2024)
von: Zhao, Michelle, et al.
Veröffentlicht: (2024)
Can We Detect Failures Without Failure Data? Uncertainty-Aware Runtime Failure Detection for Imitation Learning Policies
von: Xu, Chen, et al.
Veröffentlicht: (2025)
von: Xu, Chen, et al.
Veröffentlicht: (2025)
Robust Adversarial Policy Optimization Under Dynamics Uncertainty
von: Kim, Mintae, et al.
Veröffentlicht: (2026)
von: Kim, Mintae, et al.
Veröffentlicht: (2026)
Learning Efficient and Fair Policies for Uncertainty-Aware Collaborative Human-Robot Order Picking
von: Smit, Igor G., et al.
Veröffentlicht: (2024)
von: Smit, Igor G., et al.
Veröffentlicht: (2024)
Latent Policy Steering with Embodiment-Agnostic Pretrained World Models
von: Wang, Yiqi, et al.
Veröffentlicht: (2025)
von: Wang, Yiqi, et al.
Veröffentlicht: (2025)
Uncertainty-Aware Rank-One MIMO Q Network Framework for Accelerated Offline Reinforcement Learning
von: Nguyen, Thanh, et al.
Veröffentlicht: (2026)
von: Nguyen, Thanh, et al.
Veröffentlicht: (2026)
ActSafe: Active Exploration with Safety Constraints for Reinforcement Learning
von: As, Yarden, et al.
Veröffentlicht: (2024)
von: As, Yarden, et al.
Veröffentlicht: (2024)
Uncertainty Comes for Free: Human-in-the-Loop Policies with Diffusion Models
von: He, Zhanpeng, et al.
Veröffentlicht: (2025)
von: He, Zhanpeng, et al.
Veröffentlicht: (2025)
Uncertainty-Based Smooth Policy Regularisation for Reinforcement Learning with Few Demonstrations
von: Zhu, Yujie, et al.
Veröffentlicht: (2025)
von: Zhu, Yujie, et al.
Veröffentlicht: (2025)
Unpacking the Individual Components of Diffusion Policy
von: Yuan, Xiu
Veröffentlicht: (2024)
von: Yuan, Xiu
Veröffentlicht: (2024)
UMLoc: Uncertainty-Aware Map-Constrained Inertial Localization with Quantified Bounds
von: Alharbi, Mohammed S., et al.
Veröffentlicht: (2026)
von: Alharbi, Mohammed S., et al.
Veröffentlicht: (2026)
An Efficient Reachability-Based Framework for Provably Safe Autonomous Navigation in Unknown Environments
von: Bajcsy, Andrea, et al.
Veröffentlicht: (2019)
von: Bajcsy, Andrea, et al.
Veröffentlicht: (2019)
Learning Hybrid-Control Policies for High-Precision In-Contact Manipulation Under Uncertainty
von: Brown, Hunter L., et al.
Veröffentlicht: (2026)
von: Brown, Hunter L., et al.
Veröffentlicht: (2026)
Uncertainty-Aware Trajectory Prediction via Rule-Regularized Heteroscedastic Deep Classification
von: Manas, Kumar, et al.
Veröffentlicht: (2025)
von: Manas, Kumar, et al.
Veröffentlicht: (2025)
Active Learning of Discrete-Time Dynamics for Uncertainty-Aware Model Predictive Control
von: Saviolo, Alessandro, et al.
Veröffentlicht: (2022)
von: Saviolo, Alessandro, et al.
Veröffentlicht: (2022)
When Should a Robot Think? Resource-Aware Reasoning via Reinforcement Learning for Embodied Robotic Decision-Making
von: Liu, Jun, et al.
Veröffentlicht: (2026)
von: Liu, Jun, et al.
Veröffentlicht: (2026)
Learning to Act Robustly with View-Invariant Latent Actions
von: Jeong, Youngjoon, et al.
Veröffentlicht: (2026)
von: Jeong, Youngjoon, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
From Foresight to Forethought: VLM-In-the-Loop Policy Steering via Latent Alignment
von: Wu, Yilin, et al.
Veröffentlicht: (2025) -
Do What You Say: Steering Vision-Language-Action Models via Runtime Reasoning-Action Alignment Verification
von: Wu, Yilin, et al.
Veröffentlicht: (2025) -
Uncertainty-aware Latent Safety Filters for Avoiding Out-of-Distribution Failures
von: Seo, Junwon, et al.
Veröffentlicht: (2025) -
Maximizing Alignment with Minimal Feedback: Efficiently Learning Rewards for Visuomotor Robot Policy Alignment
von: Tian, Ran, et al.
Veröffentlicht: (2024) -
Adapting by Analogy: OOD Generalization of Visuomotor Policies via Functional Correspondence
von: Gupta, Pranay, et al.
Veröffentlicht: (2025)