CrossQ: Batch Normalization in Deep Reinforcement Learning for Greater Sample Efficiency and Simplicity
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bhatt, Aditya, Palenicek, Daniel, Belousov, Boris, Argus, Max, Amiranashvili, Artemij, Brox, Thomas, Peters, Jan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2019
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Scaling CrossQ with Weight Normalization
von: Palenicek, Daniel, et al.
Veröffentlicht: (2025)
von: Palenicek, Daniel, et al.
Veröffentlicht: (2025)
Iterated $Q$-Network: Beyond One-Step Bellman Updates in Deep Reinforcement Learning
von: Vincent, Théo, et al.
Veröffentlicht: (2024)
von: Vincent, Théo, et al.
Veröffentlicht: (2024)
Scaling Off-Policy Reinforcement Learning with Batch and Weight Normalization
von: Palenicek, Daniel, et al.
Veröffentlicht: (2025)
von: Palenicek, Daniel, et al.
Veröffentlicht: (2025)
MuJoCo MPC for Humanoid Control: Evaluation on HumanoidBench
von: Meser, Moritz, et al.
Veröffentlicht: (2024)
von: Meser, Moritz, et al.
Veröffentlicht: (2024)
Adaptive $Q$-Network: On-the-fly Target Selection for Deep Reinforcement Learning
von: Vincent, Théo, et al.
Veröffentlicht: (2024)
von: Vincent, Théo, et al.
Veröffentlicht: (2024)
When and How Does CLIP Enable Domain and Compositional Generalization?
von: Kempf, Elias, et al.
Veröffentlicht: (2025)
von: Kempf, Elias, et al.
Veröffentlicht: (2025)
Concept Bottleneck Models Without Predefined Concepts
von: Schrodi, Simon, et al.
Veröffentlicht: (2024)
von: Schrodi, Simon, et al.
Veröffentlicht: (2024)
XQC: Well-conditioned Optimization Accelerates Deep Reinforcement Learning
von: Palenicek, Daniel, et al.
Veröffentlicht: (2025)
von: Palenicek, Daniel, et al.
Veröffentlicht: (2025)
DITTO: Demonstration Imitation by Trajectory Transformation
von: Heppert, Nick, et al.
Veröffentlicht: (2024)
von: Heppert, Nick, et al.
Veröffentlicht: (2024)
Two Effects, One Trigger: On the Modality Gap, Object Bias, and Information Imbalance in Contrastive Vision-Language Models
von: Schrodi, Simon, et al.
Veröffentlicht: (2024)
von: Schrodi, Simon, et al.
Veröffentlicht: (2024)
Learning Robotic Manipulation Policies from Point Clouds with Conditional Flow Matching
von: Chisari, Eugenio, et al.
Veröffentlicht: (2024)
von: Chisari, Eugenio, et al.
Veröffentlicht: (2024)
cVLA: Towards Efficient Camera-Space VLAs
von: Argus, Max, et al.
Veröffentlicht: (2025)
von: Argus, Max, et al.
Veröffentlicht: (2025)
In-Hand Object Pose Estimation via Visual-Tactile Fusion
von: Nonnengießer, Felix, et al.
Veröffentlicht: (2025)
von: Nonnengießer, Felix, et al.
Veröffentlicht: (2025)
Analysing the Interplay of Vision and Touch for Dexterous Insertion Tasks
von: Lenz, Janis, et al.
Veröffentlicht: (2024)
von: Lenz, Janis, et al.
Veröffentlicht: (2024)
Gait in Eight: Efficient On-Robot Learning for Omnidirectional Quadruped Locomotion
von: Bohlinger, Nico, et al.
Veröffentlicht: (2025)
von: Bohlinger, Nico, et al.
Veröffentlicht: (2025)
Kaputt: A Large-Scale Dataset for Visual Defect Detection
von: Höfer, Sebastian, et al.
Veröffentlicht: (2025)
von: Höfer, Sebastian, et al.
Veröffentlicht: (2025)
DIME:Diffusion-Based Maximum Entropy Reinforcement Learning
von: Celik, Onur, et al.
Veröffentlicht: (2025)
von: Celik, Onur, et al.
Veröffentlicht: (2025)
Learning Force Distribution Estimation for the GelSight Mini Optical Tactile Sensor Based on Finite Element Analysis
von: Helmut, Erik, et al.
Veröffentlicht: (2024)
von: Helmut, Erik, et al.
Veröffentlicht: (2024)
Diminishing Return of Value Expansion Methods
von: Palenicek, Daniel, et al.
Veröffentlicht: (2024)
von: Palenicek, Daniel, et al.
Veröffentlicht: (2024)
Velocity-History-Based Soft Actor-Critic Tackling IROS'24 Competition "AI Olympics with RealAIGym"
von: Faust, Tim Lukas, et al.
Veröffentlicht: (2024)
von: Faust, Tim Lukas, et al.
Veröffentlicht: (2024)
The Role of Domain Randomization in Training Diffusion Policies for Whole-Body Humanoid Control
von: Kaidanov, Oleg, et al.
Veröffentlicht: (2024)
von: Kaidanov, Oleg, et al.
Veröffentlicht: (2024)
XQCfD: Accelerating Fast Actor-Critic Algorithms with Prior Data and Prior Policies
von: Palenicek, Daniel, et al.
Veröffentlicht: (2026)
von: Palenicek, Daniel, et al.
Veröffentlicht: (2026)
Coherent Off-Policy Improvement of Large Behavior Models with Learned Rewards
von: Scherer, Christian, et al.
Veröffentlicht: (2026)
von: Scherer, Christian, et al.
Veröffentlicht: (2026)
Sample-Efficiency in Multi-Batch Reinforcement Learning: The Need for Dimension-Dependent Adaptivity
von: Johnson, Emmeran, et al.
Veröffentlicht: (2023)
von: Johnson, Emmeran, et al.
Veröffentlicht: (2023)
Sustaining multidisciplinary teams in rural and remote primary care
von: Geoff Argus
Veröffentlicht: (2024)
von: Geoff Argus
Veröffentlicht: (2024)
Brainformers: Trading Simplicity for Efficiency
von: Zhou, Yanqi, et al.
Veröffentlicht: (2023)
von: Zhou, Yanqi, et al.
Veröffentlicht: (2023)
Eau De $Q$-Network: Adaptive Distillation of Neural Networks in Deep Reinforcement Learning
von: Vincent, Théo, et al.
Veröffentlicht: (2025)
von: Vincent, Théo, et al.
Veröffentlicht: (2025)
Smith Normal Forms of Graphical Hermite Simplices
von: Braun, Benjamin, et al.
Veröffentlicht: (2025)
von: Braun, Benjamin, et al.
Veröffentlicht: (2025)
Scalable Batch Correction for Cell Painting via Batch-Dependent Kernels and Adaptive Sampling
von: Ravi, Aditya Narayan, et al.
Veröffentlicht: (2026)
von: Ravi, Aditya Narayan, et al.
Veröffentlicht: (2026)
Making Batch Normalization Great in Federated Deep Learning
von: Zhong, Jike, et al.
Veröffentlicht: (2023)
von: Zhong, Jike, et al.
Veröffentlicht: (2023)
TacEx: GelSight Tactile Simulation in Isaac Sim -- Combining Soft-Body and Visuotactile Simulators
von: Nguyen, Duc Huy, et al.
Veröffentlicht: (2024)
von: Nguyen, Duc Huy, et al.
Veröffentlicht: (2024)
Simplicity of Augmentation Submodules in Monoids with 0-Minimal Ideals of Rank Greater than Two
von: Shahzamanian, M. H.
Veröffentlicht: (2026)
von: Shahzamanian, M. H.
Veröffentlicht: (2026)
Patch-aware Batch Normalization for Improving Cross-domain Robustness
von: Qi, Lei, et al.
Veröffentlicht: (2023)
von: Qi, Lei, et al.
Veröffentlicht: (2023)
Constrained Reinforcement Learning for Safe Heat Pump Control
von: Zhang, Baohe, et al.
Veröffentlicht: (2024)
von: Zhang, Baohe, et al.
Veröffentlicht: (2024)
Supervised Batch Normalization
von: Faye, Bilal, et al.
Veröffentlicht: (2024)
von: Faye, Bilal, et al.
Veröffentlicht: (2024)
Batch Normalization Decomposed
von: Nachum, Ido, et al.
Veröffentlicht: (2024)
von: Nachum, Ido, et al.
Veröffentlicht: (2024)
Discrete Variational Autoencoding via Policy Search
von: Drolet, Michael, et al.
Veröffentlicht: (2025)
von: Drolet, Michael, et al.
Veröffentlicht: (2025)
Parameterized Projected Bellman Operator
von: Vincent, Théo, et al.
Veröffentlicht: (2023)
von: Vincent, Théo, et al.
Veröffentlicht: (2023)
Tournaments, Contestant Heterogeneity and Performance
von: Brox, Enzo, et al.
Veröffentlicht: (2024)
von: Brox, Enzo, et al.
Veröffentlicht: (2024)
SimBa: Simplicity Bias for Scaling Up Parameters in Deep Reinforcement Learning
von: Lee, Hojoon, et al.
Veröffentlicht: (2024)
von: Lee, Hojoon, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Scaling CrossQ with Weight Normalization
von: Palenicek, Daniel, et al.
Veröffentlicht: (2025) -
Iterated $Q$-Network: Beyond One-Step Bellman Updates in Deep Reinforcement Learning
von: Vincent, Théo, et al.
Veröffentlicht: (2024) -
Scaling Off-Policy Reinforcement Learning with Batch and Weight Normalization
von: Palenicek, Daniel, et al.
Veröffentlicht: (2025) -
MuJoCo MPC for Humanoid Control: Evaluation on HumanoidBench
von: Meser, Moritz, et al.
Veröffentlicht: (2024) -
Adaptive $Q$-Network: On-the-fly Target Selection for Deep Reinforcement Learning
von: Vincent, Théo, et al.
Veröffentlicht: (2024)