Average-DICE: Stationary Distribution Correction by Regression
Fuente:
arXiv
Guardado en:
| Autores principales: | Che, Fengdi, Chan, Bryan, Ma, Chen, Mahmood, A. Rupam |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A Tutorial: An Intuitive Explanation of Offline Reinforcement Learning Theory
por: Che, Fengdi
Publicado: (2025)
por: Che, Fengdi
Publicado: (2025)
Distributions as Actions: A Unified Framework for Diverse Action Spaces
por: He, Jiamin, et al.
Publicado: (2025)
por: He, Jiamin, et al.
Publicado: (2025)
Target Networks and Over-parameterization Stabilize Off-policy Bootstrapping with Function Approximation
por: Che, Fengdi, et al.
Publicado: (2024)
por: Che, Fengdi, et al.
Publicado: (2024)
Learning Without Time-Based Embodiment Resets in Soft-Actor Critic
por: Farrahi, Homayoon, et al.
Publicado: (2025)
por: Farrahi, Homayoon, et al.
Publicado: (2025)
ComaDICE: Offline Cooperative Multi-Agent Reinforcement Learning with Stationary Distribution Shift Regularization
por: Bui, The Viet, et al.
Publicado: (2024)
por: Bui, The Viet, et al.
Publicado: (2024)
Addressing Loss of Plasticity and Catastrophic Forgetting in Continual Learning
por: Elsayed, Mohamed, et al.
Publicado: (2024)
por: Elsayed, Mohamed, et al.
Publicado: (2024)
Efficient Reinforcement Learning by Reducing Forgetting with Elephant Activation Functions
por: Lan, Qingfeng, et al.
Publicado: (2025)
por: Lan, Qingfeng, et al.
Publicado: (2025)
Streaming Deep Reinforcement Learning Finally Works
por: Elsayed, Mohamed, et al.
Publicado: (2024)
por: Elsayed, Mohamed, et al.
Publicado: (2024)
SimuDICE: Offline Policy Optimization Through World Model Updates and DICE Estimation
por: Brita, Catalin E., et al.
Publicado: (2024)
por: Brita, Catalin E., et al.
Publicado: (2024)
Robust Regression over Averaged Uncertainty
por: Bertsimas, Dimitris, et al.
Publicado: (2023)
por: Bertsimas, Dimitris, et al.
Publicado: (2023)
Revisiting Scalable Hessian Diagonal Approximations for Applications in Reinforcement Learning
por: Elsayed, Mohamed, et al.
Publicado: (2024)
por: Elsayed, Mohamed, et al.
Publicado: (2024)
Learning to Optimize for Reinforcement Learning
por: Lan, Qingfeng, et al.
Publicado: (2023)
por: Lan, Qingfeng, et al.
Publicado: (2023)
Weight Clipping for Deep Continual and Reinforcement Learning
por: Elsayed, Mohamed, et al.
Publicado: (2024)
por: Elsayed, Mohamed, et al.
Publicado: (2024)
Natural Policy Gradient for Average Reward Non-Stationary RL
por: Jali, Neharika, et al.
Publicado: (2025)
por: Jali, Neharika, et al.
Publicado: (2025)
SEMDICE: Off-policy State Entropy Maximization via Stationary Distribution Correction Estimation
por: Lee, Jongmin, et al.
Publicado: (2025)
por: Lee, Jongmin, et al.
Publicado: (2025)
Non-Stationary Latent Auto-Regressive Bandits
por: Trella, Anna L., et al.
Publicado: (2024)
por: Trella, Anna L., et al.
Publicado: (2024)
Semi-gradient DICE for Offline Constrained Reinforcement Learning
por: Kim, Woosung, et al.
Publicado: (2025)
por: Kim, Woosung, et al.
Publicado: (2025)
Maintaining Plasticity in Deep Continual Learning
por: Dohare, Shibhansh, et al.
Publicado: (2023)
por: Dohare, Shibhansh, et al.
Publicado: (2023)
ReModels: Quantile Regression Averaging models
por: Zakrzewski, Grzegorz, et al.
Publicado: (2024)
por: Zakrzewski, Grzegorz, et al.
Publicado: (2024)
DICE: Discrete inverse continuity equation for learning population dynamics
por: Blickhan, Tobias, et al.
Publicado: (2025)
por: Blickhan, Tobias, et al.
Publicado: (2025)
Intentional Updates for Streaming Reinforcement Learning
por: Sharifnassab, Arsalan, et al.
Publicado: (2026)
por: Sharifnassab, Arsalan, et al.
Publicado: (2026)
Revisiting Sparse Rewards for Goal-Reaching Reinforcement Learning
por: Vasan, Gautham, et al.
Publicado: (2024)
por: Vasan, Gautham, et al.
Publicado: (2024)
DICE: Device-level Integrated Circuits Encoder with Graph Contrastive Pretraining
por: Lee, Sungyoung, et al.
Publicado: (2025)
por: Lee, Sungyoung, et al.
Publicado: (2025)
DICE: Diffusion Large Language Models Excel at Generating CUDA Kernels
por: Bai, Haolei, et al.
Publicado: (2026)
por: Bai, Haolei, et al.
Publicado: (2026)
Provable and Practical: Efficient Exploration in Reinforcement Learning via Langevin Monte Carlo
por: Ishfaq, Haque, et al.
Publicado: (2023)
por: Ishfaq, Haque, et al.
Publicado: (2023)
Sampling from the Mean-Field Stationary Distribution
por: Kook, Yunbum, et al.
Publicado: (2024)
por: Kook, Yunbum, et al.
Publicado: (2024)
[Re] FairDICE: A Fair Tradeoff in Multi-objective Offline RL
por: Adema, Peter, et al.
Publicado: (2026)
por: Adema, Peter, et al.
Publicado: (2026)
Diffusion-DICE: In-Sample Diffusion Guidance for Offline Reinforcement Learning
por: Mao, Liyuan, et al.
Publicado: (2024)
por: Mao, Liyuan, et al.
Publicado: (2024)
Efficient Context Propagating Perceiver Architectures for Auto-Regressive Language Modeling
por: Mahmood, Kaleel, et al.
Publicado: (2024)
por: Mahmood, Kaleel, et al.
Publicado: (2024)
Node Regression on Latent Position Random Graphs via Local Averaging
por: Gjorgjevski, Martin, et al.
Publicado: (2024)
por: Gjorgjevski, Martin, et al.
Publicado: (2024)
Direct Bayesian Additive Regression Trees for Conditional Average Treatment Effects in Regression Discontinuity Designs
por: Kondo, Daisuke, et al.
Publicado: (2026)
por: Kondo, Daisuke, et al.
Publicado: (2026)
Concentration of the Langevin Algorithm's Stationary Distribution
por: Altschuler, Jason M., et al.
Publicado: (2022)
por: Altschuler, Jason M., et al.
Publicado: (2022)
Risk-averse Learning with Non-Stationary Distributions
por: Wang, Siyi, et al.
Publicado: (2024)
por: Wang, Siyi, et al.
Publicado: (2024)
FairDICE: Fairness-Driven Offline Multi-Objective Reinforcement Learning
por: Kim, Woosung, et al.
Publicado: (2025)
por: Kim, Woosung, et al.
Publicado: (2025)
OLR-WAA: Adaptive and Drift-Resilient Online Regression with Dynamic Weighted Averaging
por: Abu-Shaira, Mohammad, et al.
Publicado: (2025)
por: Abu-Shaira, Mohammad, et al.
Publicado: (2025)
Stationary MMD Points
por: Chen, Zonghao, et al.
Publicado: (2025)
por: Chen, Zonghao, et al.
Publicado: (2025)
A Simplified Analysis of SGD for Linear Regression with Weight Averaging
por: Meterez, Alexandru, et al.
Publicado: (2025)
por: Meterez, Alexandru, et al.
Publicado: (2025)
A Bayesian Additive Regression Tree Model for Learning Conditional Average Treatment Effects in Regression Discontinuity Designs
por: Alcantara, Rafael, et al.
Publicado: (2025)
por: Alcantara, Rafael, et al.
Publicado: (2025)
Noise Balance and Stationary Distribution of Stochastic Gradient Descent
por: Ziyin, Liu, et al.
Publicado: (2023)
por: Ziyin, Liu, et al.
Publicado: (2023)
Risk Bounds For Distributional Regression
por: Padilla, Carlos Misael Madrid, et al.
Publicado: (2025)
por: Padilla, Carlos Misael Madrid, et al.
Publicado: (2025)
Ejemplares similares
-
A Tutorial: An Intuitive Explanation of Offline Reinforcement Learning Theory
por: Che, Fengdi
Publicado: (2025) -
Distributions as Actions: A Unified Framework for Diverse Action Spaces
por: He, Jiamin, et al.
Publicado: (2025) -
Target Networks and Over-parameterization Stabilize Off-policy Bootstrapping with Function Approximation
por: Che, Fengdi, et al.
Publicado: (2024) -
Learning Without Time-Based Embodiment Resets in Soft-Actor Critic
por: Farrahi, Homayoon, et al.
Publicado: (2025) -
ComaDICE: Offline Cooperative Multi-Agent Reinforcement Learning with Stationary Distribution Shift Regularization
por: Bui, The Viet, et al.
Publicado: (2024)