XQCfD: Accelerating Fast Actor-Critic Algorithms with Prior Data and Prior Policies
Fuente:
arXiv
Saved in:
| Main Authors: | Palenicek, Daniel, Vogt, Florian, Watson, Joe, Posner, Ingmar, Kragic, Danica, Peters, Jan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
XQC: Well-conditioned Optimization Accelerates Deep Reinforcement Learning
by: Palenicek, Daniel, et al.
Published: (2025)
by: Palenicek, Daniel, et al.
Published: (2025)
Scaling Off-Policy Reinforcement Learning with Batch and Weight Normalization
by: Palenicek, Daniel, et al.
Published: (2025)
by: Palenicek, Daniel, et al.
Published: (2025)
Coherent Off-Policy Improvement of Large Behavior Models with Learned Rewards
by: Scherer, Christian, et al.
Published: (2026)
by: Scherer, Christian, et al.
Published: (2026)
Scaling CrossQ with Weight Normalization
by: Palenicek, Daniel, et al.
Published: (2025)
by: Palenicek, Daniel, et al.
Published: (2025)
Posterior Sampling Reinforcement Learning with Gaussian Processes for Continuous Control: Sublinear Regret Bounds for Unbounded State Spaces
by: Flynn, Hamish, et al.
Published: (2026)
by: Flynn, Hamish, et al.
Published: (2026)
FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control
by: Kim, Donghu, et al.
Published: (2026)
by: Kim, Donghu, et al.
Published: (2026)
Disentangling Dynamical Systems: Causal Representation Learning Meets Local Sparse Attention
by: Baumgartner, Markus W., et al.
Published: (2026)
by: Baumgartner, Markus W., et al.
Published: (2026)
World Models via Policy-Guided Trajectory Diffusion
by: Rigter, Marc, et al.
Published: (2023)
by: Rigter, Marc, et al.
Published: (2023)
Goal-Conditioned Reinforcement Learning from Sub-Optimal Data on Metric Spaces
by: Reichlin, Alfredo, et al.
Published: (2024)
by: Reichlin, Alfredo, et al.
Published: (2024)
Walking on the Fiber: A Simple Geometric Approximation for Bayesian Neural Networks
by: Reichlin, Alfredo, et al.
Published: (2025)
by: Reichlin, Alfredo, et al.
Published: (2025)
An Investigation of Batch Normalization in Off-Policy Actor-Critic Algorithms
by: Wang, Li, et al.
Published: (2025)
by: Wang, Li, et al.
Published: (2025)
Gait in Eight: Efficient On-Robot Learning for Omnidirectional Quadruped Locomotion
by: Bohlinger, Nico, et al.
Published: (2025)
by: Bohlinger, Nico, et al.
Published: (2025)
Real-Time Operator Takeover for Visuomotor Diffusion Policy Training
by: Moletta, Marco, et al.
Published: (2025)
by: Moletta, Marco, et al.
Published: (2025)
Diminishing Return of Value Expansion Methods
by: Palenicek, Daniel, et al.
Published: (2024)
by: Palenicek, Daniel, et al.
Published: (2024)
Learning Hamiltonian Dynamics at Scale: A Differential-Geometric Approach
by: Friedl, Katharina, et al.
Published: (2025)
by: Friedl, Katharina, et al.
Published: (2025)
Geometry of Uncertainty: Learning Metric Spaces for Multimodal State Estimation in RL
by: Reichlin, Alfredo, et al.
Published: (2026)
by: Reichlin, Alfredo, et al.
Published: (2026)
Iterated $Q$-Network: Beyond One-Step Bellman Updates in Deep Reinforcement Learning
by: Vincent, Théo, et al.
Published: (2024)
by: Vincent, Théo, et al.
Published: (2024)
Reward-Free Curricula for Training Robust World Models
by: Rigter, Marc, et al.
Published: (2023)
by: Rigter, Marc, et al.
Published: (2023)
SPARTAN: A Sparse Transformer World Model Attending to What Matters
by: Lei, Anson, et al.
Published: (2024)
by: Lei, Anson, et al.
Published: (2024)
Limits of Actor-Critic Algorithms for Decision Tree Policies Learning in IBMDPs
by: Kohler, Hector, et al.
Published: (2023)
by: Kohler, Hector, et al.
Published: (2023)
DexDiffuser: Generating Dexterous Grasps with Diffusion Models
by: Weng, Zehang, et al.
Published: (2024)
by: Weng, Zehang, et al.
Published: (2024)
D2 Actor Critic: Diffusion Actor Meets Distributional Critic
by: Zhang, Lunjun, et al.
Published: (2025)
by: Zhang, Lunjun, et al.
Published: (2025)
D-Cubed: Latent Diffusion Trajectory Optimisation for Dexterous Deformable Manipulation
by: Yamada, Jun, et al.
Published: (2024)
by: Yamada, Jun, et al.
Published: (2024)
Grasping a Handful: Sequential Multi-Object Dexterous Grasp Generation
by: Lu, Haofei, et al.
Published: (2025)
by: Lu, Haofei, et al.
Published: (2025)
POP: Prior-Fitted First-Order Optimization Policies
by: Kobiolka, Jan, et al.
Published: (2026)
by: Kobiolka, Jan, et al.
Published: (2026)
Actor-Critic Pretraining for Proximal Policy Optimization
by: Kernbach, Andreas, et al.
Published: (2026)
by: Kernbach, Andreas, et al.
Published: (2026)
Compatible Gradient Approximations for Actor-Critic Algorithms
by: Saglam, Baturay, et al.
Published: (2024)
by: Saglam, Baturay, et al.
Published: (2024)
Finite-Time Analysis of Three-Timescale Constrained Actor-Critic and Constrained Natural Actor-Critic Algorithms
by: Panda, Prashansa, et al.
Published: (2023)
by: Panda, Prashansa, et al.
Published: (2023)
Value Improved Actor Critic Algorithms
by: Oren, Yaniv, et al.
Published: (2024)
by: Oren, Yaniv, et al.
Published: (2024)
WarmPrior: Straightening Flow-Matching Policies with Temporal Priors
by: Kang, Sinjae, et al.
Published: (2026)
by: Kang, Sinjae, et al.
Published: (2026)
CrossQ: Batch Normalization in Deep Reinforcement Learning for Greater Sample Efficiency and Simplicity
by: Bhatt, Aditya, et al.
Published: (2019)
by: Bhatt, Aditya, et al.
Published: (2019)
Actor-Critic without Actor
by: Ki, Donghyeon, et al.
Published: (2025)
by: Ki, Donghyeon, et al.
Published: (2025)
Relative Representations: Topological and Geometric Perspectives
by: García-Castellanos, Alejandro, et al.
Published: (2024)
by: García-Castellanos, Alejandro, et al.
Published: (2024)
A Riemannian Framework for Learning Reduced-order Lagrangian Dynamics
by: Friedl, Katharina, et al.
Published: (2024)
by: Friedl, Katharina, et al.
Published: (2024)
Simulation Priors for Data-Efficient Deep Learning
by: Treven, Lenart, et al.
Published: (2025)
by: Treven, Lenart, et al.
Published: (2025)
DeepRV: Accelerating Spatiotemporal Inference with Pre-trained Neural Priors
by: Navott, Jhonathan, et al.
Published: (2025)
by: Navott, Jhonathan, et al.
Published: (2025)
Actor-Critic Algorithm for Dynamic Expectile and CVaR
by: Luo, Yudong, et al.
Published: (2026)
by: Luo, Yudong, et al.
Published: (2026)
A Theoretical Justification for Asymmetric Actor-Critic Algorithms
by: Lambrechts, Gaspard, et al.
Published: (2025)
by: Lambrechts, Gaspard, et al.
Published: (2025)
Generative Actor-Critic with Soft Bridge Policies
by: He, Ke, et al.
Published: (2026)
by: He, Ke, et al.
Published: (2026)
Distributional Soft Actor-Critic with Diffusion Policy
by: Liu, Tong, et al.
Published: (2025)
by: Liu, Tong, et al.
Published: (2025)
Similar Items
-
XQC: Well-conditioned Optimization Accelerates Deep Reinforcement Learning
by: Palenicek, Daniel, et al.
Published: (2025) -
Scaling Off-Policy Reinforcement Learning with Batch and Weight Normalization
by: Palenicek, Daniel, et al.
Published: (2025) -
Coherent Off-Policy Improvement of Large Behavior Models with Learned Rewards
by: Scherer, Christian, et al.
Published: (2026) -
Scaling CrossQ with Weight Normalization
by: Palenicek, Daniel, et al.
Published: (2025) -
Posterior Sampling Reinforcement Learning with Gaussian Processes for Continuous Control: Sublinear Regret Bounds for Unbounded State Spaces
by: Flynn, Hamish, et al.
Published: (2026)