Average-DICE: Stationary Distribution Correction by Regression
Fuente:
arXiv
Saved in:
| Main Authors: | Che, Fengdi, Chan, Bryan, Ma, Chen, Mahmood, A. Rupam |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Tutorial: An Intuitive Explanation of Offline Reinforcement Learning Theory
by: Che, Fengdi
Published: (2025)
by: Che, Fengdi
Published: (2025)
Distributions as Actions: A Unified Framework for Diverse Action Spaces
by: He, Jiamin, et al.
Published: (2025)
by: He, Jiamin, et al.
Published: (2025)
Target Networks and Over-parameterization Stabilize Off-policy Bootstrapping with Function Approximation
by: Che, Fengdi, et al.
Published: (2024)
by: Che, Fengdi, et al.
Published: (2024)
Learning Without Time-Based Embodiment Resets in Soft-Actor Critic
by: Farrahi, Homayoon, et al.
Published: (2025)
by: Farrahi, Homayoon, et al.
Published: (2025)
ComaDICE: Offline Cooperative Multi-Agent Reinforcement Learning with Stationary Distribution Shift Regularization
by: Bui, The Viet, et al.
Published: (2024)
by: Bui, The Viet, et al.
Published: (2024)
Addressing Loss of Plasticity and Catastrophic Forgetting in Continual Learning
by: Elsayed, Mohamed, et al.
Published: (2024)
by: Elsayed, Mohamed, et al.
Published: (2024)
Efficient Reinforcement Learning by Reducing Forgetting with Elephant Activation Functions
by: Lan, Qingfeng, et al.
Published: (2025)
by: Lan, Qingfeng, et al.
Published: (2025)
Streaming Deep Reinforcement Learning Finally Works
by: Elsayed, Mohamed, et al.
Published: (2024)
by: Elsayed, Mohamed, et al.
Published: (2024)
SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation
by: Brita, Catalin E., et al.
Published: (2024)
by: Brita, Catalin E., et al.
Published: (2024)
Robust Regression over Averaged Uncertainty
by: Bertsimas, Dimitris, et al.
Published: (2023)
by: Bertsimas, Dimitris, et al.
Published: (2023)
Revisiting Scalable Hessian Diagonal Approximations for Applications in Reinforcement Learning
by: Elsayed, Mohamed, et al.
Published: (2024)
by: Elsayed, Mohamed, et al.
Published: (2024)
Learning to Optimize for Reinforcement Learning
by: Lan, Qingfeng, et al.
Published: (2023)
by: Lan, Qingfeng, et al.
Published: (2023)
Weight Clipping for Deep Continual and Reinforcement Learning
by: Elsayed, Mohamed, et al.
Published: (2024)
by: Elsayed, Mohamed, et al.
Published: (2024)
Natural Policy Gradient for Average Reward Non-Stationary RL
by: Jali, Neharika, et al.
Published: (2025)
by: Jali, Neharika, et al.
Published: (2025)
SEMDICE: Off-policy State Entropy Maximization via Stationary Distribution Correction Estimation
by: Lee, Jongmin, et al.
Published: (2025)
by: Lee, Jongmin, et al.
Published: (2025)
Non-Stationary Latent Auto-Regressive Bandits
by: Trella, Anna L., et al.
Published: (2024)
by: Trella, Anna L., et al.
Published: (2024)
Semi-gradient DICE for Offline Constrained Reinforcement Learning
by: Kim, Woosung, et al.
Published: (2025)
by: Kim, Woosung, et al.
Published: (2025)
Maintaining Plasticity in Deep Continual Learning
by: Dohare, Shibhansh, et al.
Published: (2023)
by: Dohare, Shibhansh, et al.
Published: (2023)
ReModels: Quantile Regression Averaging models
by: Zakrzewski, Grzegorz, et al.
Published: (2024)
by: Zakrzewski, Grzegorz, et al.
Published: (2024)
DICE: Discrete inverse continuity equation for learning population dynamics
by: Blickhan, Tobias, et al.
Published: (2025)
by: Blickhan, Tobias, et al.
Published: (2025)
Intentional Updates for Streaming Reinforcement Learning
by: Sharifnassab, Arsalan, et al.
Published: (2026)
by: Sharifnassab, Arsalan, et al.
Published: (2026)
Revisiting Sparse Rewards for Goal-Reaching Reinforcement Learning
by: Vasan, Gautham, et al.
Published: (2024)
by: Vasan, Gautham, et al.
Published: (2024)
DICE: Device-level Integrated Circuits Encoder with Graph Contrastive Pretraining
by: Lee, Sungyoung, et al.
Published: (2025)
by: Lee, Sungyoung, et al.
Published: (2025)
DICE: Diffusion Large Language Models Excel at Generating CUDA Kernels
by: Bai, Haolei, et al.
Published: (2026)
by: Bai, Haolei, et al.
Published: (2026)
Provable and Practical: Efficient Exploration in Reinforcement Learning via Langevin Monte Carlo
by: Ishfaq, Haque, et al.
Published: (2023)
by: Ishfaq, Haque, et al.
Published: (2023)
Sampling from the Mean-Field Stationary Distribution
by: Kook, Yunbum, et al.
Published: (2024)
by: Kook, Yunbum, et al.
Published: (2024)
[Re] FairDICE: A Fair Tradeoff in Multi-objective Offline RL
by: Adema, Peter, et al.
Published: (2026)
by: Adema, Peter, et al.
Published: (2026)
Diffusion-DICE: In-Sample Diffusion Guidance for Offline Reinforcement Learning
by: Mao, Liyuan, et al.
Published: (2024)
by: Mao, Liyuan, et al.
Published: (2024)
Efficient Context Propagating Perceiver Architectures for Auto-Regressive Language Modeling
by: Mahmood, Kaleel, et al.
Published: (2024)
by: Mahmood, Kaleel, et al.
Published: (2024)
Node Regression on Latent Position Random Graphs via Local Averaging
by: Gjorgjevski, Martin, et al.
Published: (2024)
by: Gjorgjevski, Martin, et al.
Published: (2024)
Direct Bayesian Additive Regression Trees for Conditional Average Treatment Effects in Regression Discontinuity Designs
by: Kondo, Daisuke, et al.
Published: (2026)
by: Kondo, Daisuke, et al.
Published: (2026)
Concentration of the Langevin Algorithm's Stationary Distribution
by: Altschuler, Jason M., et al.
Published: (2022)
by: Altschuler, Jason M., et al.
Published: (2022)
Risk-averse Learning with Non-Stationary Distributions
by: Wang, Siyi, et al.
Published: (2024)
by: Wang, Siyi, et al.
Published: (2024)
FairDICE: Fairness-Driven Offline Multi-Objective Reinforcement Learning
by: Kim, Woosung, et al.
Published: (2025)
by: Kim, Woosung, et al.
Published: (2025)
OLR-WAA: Adaptive and Drift-Resilient Online Regression with Dynamic Weighted Averaging
by: Abu-Shaira, Mohammad, et al.
Published: (2025)
by: Abu-Shaira, Mohammad, et al.
Published: (2025)
Stationary MMD Points
by: Chen, Zonghao, et al.
Published: (2025)
by: Chen, Zonghao, et al.
Published: (2025)
A Simplified Analysis of SGD for Linear Regression with Weight Averaging
by: Meterez, Alexandru, et al.
Published: (2025)
by: Meterez, Alexandru, et al.
Published: (2025)
A Bayesian Additive Regression Tree Model for Learning Conditional Average Treatment Effects in Regression Discontinuity Designs
by: Alcantara, Rafael, et al.
Published: (2025)
by: Alcantara, Rafael, et al.
Published: (2025)
Noise Balance and Stationary Distribution of Stochastic Gradient Descent
by: Ziyin, Liu, et al.
Published: (2023)
by: Ziyin, Liu, et al.
Published: (2023)
Risk Bounds For Distributional Regression
by: Padilla, Carlos Misael Madrid, et al.
Published: (2025)
by: Padilla, Carlos Misael Madrid, et al.
Published: (2025)
Similar Items
-
A Tutorial: An Intuitive Explanation of Offline Reinforcement Learning Theory
by: Che, Fengdi
Published: (2025) -
Distributions as Actions: A Unified Framework for Diverse Action Spaces
by: He, Jiamin, et al.
Published: (2025) -
Target Networks and Over-parameterization Stabilize Off-policy Bootstrapping with Function Approximation
by: Che, Fengdi, et al.
Published: (2024) -
Learning Without Time-Based Embodiment Resets in Soft-Actor Critic
by: Farrahi, Homayoon, et al.
Published: (2025) -
ComaDICE: Offline Cooperative Multi-Agent Reinforcement Learning with Stationary Distribution Shift Regularization
by: Bui, The Viet, et al.
Published: (2024)