Saved in:
| Main Authors: | Byun, Ju-Seung, Perrault, Andrew |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2208.13125 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Symmetric Reinforcement Learning Loss for Robust Learning on Diverse Tasks and Model Scales
by: Byun, Ju-Seung, et al.
Published: (2024)
by: Byun, Ju-Seung, et al.
Published: (2024)
DLPO: Diffusion Model Loss-Guided Reinforcement Learning for Fine-Tuning Text-to-Speech Diffusion Models
by: Chen, Jingyi, et al.
Published: (2024)
by: Chen, Jingyi, et al.
Published: (2024)
ARES: Alternating Reinforcement Learning and Supervised Fine-Tuning for Enhanced Multi-Modal Chain-of-Thought Reasoning Through Diverse AI Feedback
by: Byun, Ju-Seung, et al.
Published: (2024)
by: Byun, Ju-Seung, et al.
Published: (2024)
Fine-Tuning Text-to-Speech Diffusion Models Using Reinforcement Learning with Human Feedback
by: Chen, Jingyi, et al.
Published: (2025)
by: Chen, Jingyi, et al.
Published: (2025)
Recovering Physical Dynamics from Discrete Observations via Intrinsic Differential Consistency
by: Luo, Yuxiang, et al.
Published: (2026)
by: Luo, Yuxiang, et al.
Published: (2026)
Continuous Control Reinforcement Learning: Distributed Distributional DrQ Algorithms
by: Zhou, Zehao
Published: (2024)
by: Zhou, Zehao
Published: (2024)
Operator-Guided Invariance Learning for Continuous Reinforcement Learning
by: Zhang, Zuyuan, et al.
Published: (2026)
by: Zhang, Zuyuan, et al.
Published: (2026)
Leaving the Nest: Going Beyond Local Loss Functions for Predict-Then-Optimize
by: Shah, Sanket, et al.
Published: (2023)
by: Shah, Sanket, et al.
Published: (2023)
Test-driven Reinforcement Learning in Continuous Control
by: Yu, Zhao, et al.
Published: (2025)
by: Yu, Zhao, et al.
Published: (2025)
Optimizing Urban Service Allocation with Time-Constrained Restless Bandits
by: Mao, Yi, et al.
Published: (2025)
by: Mao, Yi, et al.
Published: (2025)
Self-Normalized Resets for Plasticity in Continual Learning
by: Farias, Vivek F., et al.
Published: (2024)
by: Farias, Vivek F., et al.
Published: (2024)
Towards Batch-to-Streaming Deep Reinforcement Learning for Continuous Control
by: De Monte, Riccardo, et al.
Published: (2026)
by: De Monte, Riccardo, et al.
Published: (2026)
Broad Critic Deep Actor Reinforcement Learning for Continuous Control
by: Thalagala, Shiron, et al.
Published: (2024)
by: Thalagala, Shiron, et al.
Published: (2024)
Regret-Guided Search Control for Efficient Learning in AlphaZero
by: Tsai, Yun-Jui, et al.
Published: (2026)
by: Tsai, Yun-Jui, et al.
Published: (2026)
CLeAN: Continual Learning Adaptive Normalization in Dynamic Environments
by: Marasco, Isabella, et al.
Published: (2026)
by: Marasco, Isabella, et al.
Published: (2026)
Discretizing Continuous Action Space with Unimodal Probability Distributions for On-Policy Reinforcement Learning
by: Zhu, Yuanyang, et al.
Published: (2024)
by: Zhu, Yuanyang, et al.
Published: (2024)
Bridging Distribution Gaps in Time Series Foundation Model Pretraining with Prototype-Guided Normalization
by: Gong, Peiliang, et al.
Published: (2025)
by: Gong, Peiliang, et al.
Published: (2025)
Bellman Unbiasedness: Toward Provably Efficient Distributional Reinforcement Learning with General Value Function Approximation
by: Cho, Taehyun, et al.
Published: (2024)
by: Cho, Taehyun, et al.
Published: (2024)
Reinforcement Learning with Euclidean Data Augmentation for State-Based Continuous Control
by: Luo, Jinzhu, et al.
Published: (2024)
by: Luo, Jinzhu, et al.
Published: (2024)
Towards Interpretable Reinforcement Learning with Constrained Normalizing Flow Policies
by: Rietz, Finn, et al.
Published: (2024)
by: Rietz, Finn, et al.
Published: (2024)
Scaling Off-Policy Reinforcement Learning with Batch and Weight Normalization
by: Palenicek, Daniel, et al.
Published: (2025)
by: Palenicek, Daniel, et al.
Published: (2025)
Opinion-Guided Reinforcement Learning
by: Dagenais, Kyanna, et al.
Published: (2024)
by: Dagenais, Kyanna, et al.
Published: (2024)
The Cell Must Go On: Agar.io for Continual Reinforcement Learning
by: Mohamed, Mohamed A., et al.
Published: (2025)
by: Mohamed, Mohamed A., et al.
Published: (2025)
Monte Carlo Beam Search for Actor-Critic Reinforcement Learning in Continuous Control
by: Alzorgan, Hazim, et al.
Published: (2025)
by: Alzorgan, Hazim, et al.
Published: (2025)
Reflex: Reinforcement Learning with Reflection Symmetry Exploitation in State-Based Continuous Control
by: Zhen, Shuai, et al.
Published: (2026)
by: Zhen, Shuai, et al.
Published: (2026)
Personalized Observation Normalization for Federated Reinforcement Learning in Simulation Environments with Heterogeneity
by: Pang, Yiran, et al.
Published: (2026)
by: Pang, Yiran, et al.
Published: (2026)
Reinforcement Learning-Guided Semi-Supervised Learning
by: Heidari, Marzi, et al.
Published: (2024)
by: Heidari, Marzi, et al.
Published: (2024)
Reinforcement Learning by Guided Safe Exploration
by: Yang, Qisong, et al.
Published: (2023)
by: Yang, Qisong, et al.
Published: (2023)
SymCircuit: Bayesian Structure Inference for Tractable Probabilistic Circuits via Entropy-Regularized Reinforcement Learning
by: Ju, Y. Sungtaek
Published: (2026)
by: Ju, Y. Sungtaek
Published: (2026)
A New Perspective on Transformers in Online Reinforcement Learning for Continuous Control
by: Kachaev, Nikita, et al.
Published: (2025)
by: Kachaev, Nikita, et al.
Published: (2025)
Don't Forget the Critic: Value-Based Data Rehearsal for Multi-Cyclic Continual Reinforcement Learning
by: Poole, Benjamin, et al.
Published: (2026)
by: Poole, Benjamin, et al.
Published: (2026)
Continual Learning as Computationally Constrained Reinforcement Learning
by: Kumar, Saurabh, et al.
Published: (2023)
by: Kumar, Saurabh, et al.
Published: (2023)
Network Distributed Multi-Agent Reinforcement Learning for Consensus Control of Quadcopters
by: Mahran, Youssef, et al.
Published: (2026)
by: Mahran, Youssef, et al.
Published: (2026)
Overcoming Slow Decision Frequencies in Continuous Control: Model-Based Sequence Reinforcement Learning for Model-Free Control
by: Patel, Devdhar, et al.
Published: (2024)
by: Patel, Devdhar, et al.
Published: (2024)
Rethinking the Foundations for Continual Reinforcement Learning
by: Elelimy, Esraa, et al.
Published: (2025)
by: Elelimy, Esraa, et al.
Published: (2025)
A Survey of Continual Reinforcement Learning
by: Pan, Chaofan, et al.
Published: (2025)
by: Pan, Chaofan, et al.
Published: (2025)
Parseval Regularization for Continual Reinforcement Learning
by: Chung, Wesley, et al.
Published: (2024)
by: Chung, Wesley, et al.
Published: (2024)
Discrete Flow Matching for Offline-to-Online Reinforcement Learning
by: Khan, Fairoz Nower, et al.
Published: (2026)
by: Khan, Fairoz Nower, et al.
Published: (2026)
Learning-based Autonomous Oversteer Control and Collision Avoidance
by: Lee, Seokjun, et al.
Published: (2025)
by: Lee, Seokjun, et al.
Published: (2025)
Demonstration Guided Multi-Objective Reinforcement Learning
by: Lu, Junlin, et al.
Published: (2024)
by: Lu, Junlin, et al.
Published: (2024)
Similar Items
-
Symmetric Reinforcement Learning Loss for Robust Learning on Diverse Tasks and Model Scales
by: Byun, Ju-Seung, et al.
Published: (2024) -
DLPO: Diffusion Model Loss-Guided Reinforcement Learning for Fine-Tuning Text-to-Speech Diffusion Models
by: Chen, Jingyi, et al.
Published: (2024) -
ARES: Alternating Reinforcement Learning and Supervised Fine-Tuning for Enhanced Multi-Modal Chain-of-Thought Reasoning Through Diverse AI Feedback
by: Byun, Ju-Seung, et al.
Published: (2024) -
Fine-Tuning Text-to-Speech Diffusion Models Using Reinforcement Learning with Human Feedback
by: Chen, Jingyi, et al.
Published: (2025) -
Recovering Physical Dynamics from Discrete Observations via Intrinsic Differential Consistency
by: Luo, Yuxiang, et al.
Published: (2026)