Q-Guided Stein Variational Model Predictive Control via RL-informed Policy Prior
Fuente:
arXiv
Saved in:
| Main Authors: | Cai, Shizhe, Yin, Zeya, Jacob, Jayadeep, Ramos, Fabio |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Deformable Cluster Manipulation via Whole-Arm Policy Learning
by: Jacob, Jayadeep, et al.
Published: (2025)
by: Jacob, Jayadeep, et al.
Published: (2025)
Safe Exploration via Policy Priors
by: Wendl, Manuel, et al.
Published: (2026)
by: Wendl, Manuel, et al.
Published: (2026)
Harnessing Bounded-Support Evolution Strategies for Policy Refinement
by: Hirschowitz, Ethan, et al.
Published: (2025)
by: Hirschowitz, Ethan, et al.
Published: (2025)
FlowQ: Energy-Guided Flow Policies for Offline Reinforcement Learning
by: Alles, Marvin, et al.
Published: (2025)
by: Alles, Marvin, et al.
Published: (2025)
A Review of Online Diffusion Policy RL Algorithms for Scalable Robotic Control
by: Choi, Wonhyeok, et al.
Published: (2026)
by: Choi, Wonhyeok, et al.
Published: (2026)
Discrete Variational Autoencoding via Policy Search
by: Drolet, Michael, et al.
Published: (2025)
by: Drolet, Michael, et al.
Published: (2025)
Grasp-MPC: Closed-Loop Visual Grasping via Value-Guided Model Predictive Control
by: Yamada, Jun, et al.
Published: (2025)
by: Yamada, Jun, et al.
Published: (2025)
WarmPrior: Straightening Flow-Matching Policies with Temporal Priors
by: Kang, Sinjae, et al.
Published: (2026)
by: Kang, Sinjae, et al.
Published: (2026)
PEEK: Guiding and Minimal Image Representations for Zero-Shot Generalization of Robot Manipulation Policies
by: Zhang, Jesse, et al.
Published: (2025)
by: Zhang, Jesse, et al.
Published: (2025)
A Multi-Fidelity Control Variate Approach for Policy Gradient Estimation
by: Liu, Xinjie, et al.
Published: (2025)
by: Liu, Xinjie, et al.
Published: (2025)
Mobile-TeleVision: Predictive Motion Priors for Humanoid Whole-Body Control
by: Lu, Chenhao, et al.
Published: (2024)
by: Lu, Chenhao, et al.
Published: (2024)
CaRL: Learning Scalable Planning Policies with Simple Rewards
by: Jaeger, Bernhard, et al.
Published: (2025)
by: Jaeger, Bernhard, et al.
Published: (2025)
Dual Online Stein Variational Inference for Control and Dynamics
by: Barcelos, Lucas, et al.
Published: (2021)
by: Barcelos, Lucas, et al.
Published: (2021)
Posterior Behavioral Cloning: Pretraining BC Policies for Efficient RL Finetuning
by: Wagenmaker, Andrew, et al.
Published: (2025)
by: Wagenmaker, Andrew, et al.
Published: (2025)
Policy-Guided Diffusion
by: Jackson, Matthew Thomas, et al.
Published: (2024)
by: Jackson, Matthew Thomas, et al.
Published: (2024)
Actor-Free Continuous Control via Structurally Maximizable Q-Functions
by: Korkmaz, Yigit, et al.
Published: (2025)
by: Korkmaz, Yigit, et al.
Published: (2025)
Efficient Online RL Fine Tuning with Offline Pre-trained Policy Only
by: Xiao, Wei, et al.
Published: (2025)
by: Xiao, Wei, et al.
Published: (2025)
Variational Distillation of Diffusion Policies into Mixture of Experts
by: Zhou, Hongyi, et al.
Published: (2024)
by: Zhou, Hongyi, et al.
Published: (2024)
MaxInfoRL: Boosting exploration in reinforcement learning through information gain maximization
by: Sukhija, Bhavya, et al.
Published: (2024)
by: Sukhija, Bhavya, et al.
Published: (2024)
Mitigating Suboptimality of Deterministic Policy Gradients in Complex Q-functions
by: Jain, Ayush, et al.
Published: (2024)
by: Jain, Ayush, et al.
Published: (2024)
QMP: Q-switch Mixture of Policies for Multi-Task Behavior Sharing
by: Zhang, Grace, et al.
Published: (2023)
by: Zhang, Grace, et al.
Published: (2023)
Bootstrapped Model Predictive Control
by: Wang, Yuhang, et al.
Published: (2025)
by: Wang, Yuhang, et al.
Published: (2025)
SPAARS: Safer RL Policy Alignment through Abstract Exploration and Refined Exploitation of Action Space
by: K, Swaminathan S, et al.
Published: (2026)
by: K, Swaminathan S, et al.
Published: (2026)
Learning Long-Context Diffusion Policies via Past-Token Prediction
by: Torne, Marcel, et al.
Published: (2025)
by: Torne, Marcel, et al.
Published: (2025)
GRAM: Generalization in Deep RL with a Robust Adaptation Module
by: Queeney, James, et al.
Published: (2024)
by: Queeney, James, et al.
Published: (2024)
Guided Policy Optimization under Partial Observability
by: Li, Yueheng, et al.
Published: (2025)
by: Li, Yueheng, et al.
Published: (2025)
Watch Less, Feel More: Sim-to-Real RL for Generalizable Articulated Object Manipulation via Motion Adaptation and Impedance Control
by: Do, Tan-Dzung, et al.
Published: (2025)
by: Do, Tan-Dzung, et al.
Published: (2025)
Growing Q-Networks: Solving Continuous Control Tasks with Adaptive Control Resolution
by: Seyde, Tim, et al.
Published: (2024)
by: Seyde, Tim, et al.
Published: (2024)
CtRL-Sim: Reactive and Controllable Driving Agents with Offline Reinforcement Learning
by: Rowe, Luke, et al.
Published: (2024)
by: Rowe, Luke, et al.
Published: (2024)
Investigating Memory in Model-Free RL with POPGym Arcade
by: Wang, Zekang, et al.
Published: (2025)
by: Wang, Zekang, et al.
Published: (2025)
DISC: Decoupling Instruction from State-Conditioned Control via Policy Generation
by: Ren, Hanxiang, et al.
Published: (2026)
by: Ren, Hanxiang, et al.
Published: (2026)
DreamControl: Human-Inspired Whole-Body Humanoid Control for Scene Interaction via Guided Diffusion
by: Kalaria, Dvij, et al.
Published: (2025)
by: Kalaria, Dvij, et al.
Published: (2025)
Align and Filter: Improving Performance in Asynchronous On-Policy RL
by: Honari, Homayoun, et al.
Published: (2026)
by: Honari, Homayoun, et al.
Published: (2026)
Do What You Say: Steering Vision-Language-Action Models via Runtime Reasoning-Action Alignment Verification
by: Wu, Yilin, et al.
Published: (2025)
by: Wu, Yilin, et al.
Published: (2025)
First Order Model-Based RL through Decoupled Backpropagation
by: Amigo, Joseph, et al.
Published: (2025)
by: Amigo, Joseph, et al.
Published: (2025)
Adaptive Advantage-Guided Policy Regularization for Offline Reinforcement Learning
by: Liu, Tenglong, et al.
Published: (2024)
by: Liu, Tenglong, et al.
Published: (2024)
Continuous Control Reinforcement Learning: Distributed Distributional DrQ Algorithms
by: Zhou, Zehao
Published: (2024)
by: Zhou, Zehao
Published: (2024)
ReFORM: Reflected Flows for On-support Offline RL via Noise Manipulation
by: Zhang, Songyuan, et al.
Published: (2026)
by: Zhang, Songyuan, et al.
Published: (2026)
ESCORT: Efficient Stein-variational and Sliced Consistency-Optimized Temporal Belief Representation for POMDPs
by: Zhang, Yunuo, et al.
Published: (2025)
by: Zhang, Yunuo, et al.
Published: (2025)
Generalised Linear Models in Deep Bayesian RL with Learnable Basis Functions
by: You, Jingyang, et al.
Published: (2025)
by: You, Jingyang, et al.
Published: (2025)
Similar Items
-
Deformable Cluster Manipulation via Whole-Arm Policy Learning
by: Jacob, Jayadeep, et al.
Published: (2025) -
Safe Exploration via Policy Priors
by: Wendl, Manuel, et al.
Published: (2026) -
Harnessing Bounded-Support Evolution Strategies for Policy Refinement
by: Hirschowitz, Ethan, et al.
Published: (2025) -
FlowQ: Energy-Guided Flow Policies for Offline Reinforcement Learning
by: Alles, Marvin, et al.
Published: (2025) -
A Review of Online Diffusion Policy RL Algorithms for Scalable Robotic Control
by: Choi, Wonhyeok, et al.
Published: (2026)