The Central Role of the Loss Function in Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Kaiwen, Kallus, Nathan, Sun, Wen |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Efficient Inference for Inverse Reinforcement Learning and Dynamic Discrete Choice Models
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
Automatic Debiased Machine Learning for Smooth Functionals of Nonparametric M-Estimands
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
Demistifying Inference after Adaptive Experiments
by: Bibaut, Aurélien, et al.
Published: (2024)
by: Bibaut, Aurélien, et al.
Published: (2024)
Exploration in the Limit
by: Cho, Brian M., et al.
Published: (2025)
by: Cho, Brian M., et al.
Published: (2025)
Clustered Switchback Designs for Experimentation Under Spatio-temporal Interference
by: Jia, Su, et al.
Published: (2023)
by: Jia, Su, et al.
Published: (2023)
Near-Optimal Non-Parametric Sequential Tests and Confidence Sequences with Possibly Dependent Observations
by: Bibaut, Aurelien, et al.
Published: (2022)
by: Bibaut, Aurelien, et al.
Published: (2022)
Nonparametric Instrumental Variable Inference with Many Weak Instruments
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
Nonparametric Instrumental Variable Analysis Without Structural Equations: Debiased Inference on Functionals of Inverse Problems with No Solutions
by: Shen, Zikai, et al.
Published: (2026)
by: Shen, Zikai, et al.
Published: (2026)
A Reductions Approach to Risk-Sensitive Reinforcement Learning with Optimized Certainty Equivalents
by: Wang, Kaiwen, et al.
Published: (2024)
by: Wang, Kaiwen, et al.
Published: (2024)
Efficient Inference after Directionally Stable Adaptive Experiments
by: Shen, Zikai, et al.
Published: (2026)
by: Shen, Zikai, et al.
Published: (2026)
Ratio-based Loss Functions
by: Helgerth, Lena, et al.
Published: (2026)
by: Helgerth, Lena, et al.
Published: (2026)
Smooth Non-Stationary Bandits
by: Jia, Su, et al.
Published: (2023)
by: Jia, Su, et al.
Published: (2023)
More Benefits of Being Distributional: Second-Order Bounds for Reinforcement Learning
by: Wang, Kaiwen, et al.
Published: (2024)
by: Wang, Kaiwen, et al.
Published: (2024)
Centrality Estimators for Probability Density Functions
by: Ziou, Djemel
Published: (2024)
by: Ziou, Djemel
Published: (2024)
Provable Reward-Agnostic Preference-Based Reinforcement Learning
by: Zhan, Wenhao, et al.
Published: (2023)
by: Zhan, Wenhao, et al.
Published: (2023)
Federated Optimization of Smooth Loss Functions
by: Jadbabaie, Ali, et al.
Published: (2022)
by: Jadbabaie, Ali, et al.
Published: (2022)
Learning from Aggregate responses: Instance Level versus Bag Level Loss Functions
by: Javanmard, Adel, et al.
Published: (2024)
by: Javanmard, Adel, et al.
Published: (2024)
Pseudo-Labeling for Unsupervised Domain Adaptation with Kernel GLMs
by: Weill, Nathan, et al.
Published: (2026)
by: Weill, Nathan, et al.
Published: (2026)
Learning single-index models via harmonic decomposition
by: Joshi, Nirmit, et al.
Published: (2025)
by: Joshi, Nirmit, et al.
Published: (2025)
Learning from Samples: Inverse Problems over measures via Sharpened Fenchel-Young Losses
by: Andrade, Francisco, et al.
Published: (2025)
by: Andrade, Francisco, et al.
Published: (2025)
Statistical analysis of Inverse Entropy-regularized Reinforcement Learning
by: Belomestny, Denis, et al.
Published: (2025)
by: Belomestny, Denis, et al.
Published: (2025)
On the Role of Depth and Looping for In-Context Learning with Task Diversity
by: Gatmiry, Khashayar, et al.
Published: (2024)
by: Gatmiry, Khashayar, et al.
Published: (2024)
Anytime-Valid Continuous-Time Confidence Processes for Inhomogeneous Poisson Processes
by: Lindon, Michael, et al.
Published: (2024)
by: Lindon, Michael, et al.
Published: (2024)
Statistical Learning Guarantees for Group-Invariant Barron Functions
by: Yang, Yahong, et al.
Published: (2025)
by: Yang, Yahong, et al.
Published: (2025)
Decomposing Probabilistic Scores: Reliability, Information Loss and Uncertainty
by: Charpentier, Arthur, et al.
Published: (2026)
by: Charpentier, Arthur, et al.
Published: (2026)
A Unified Information-Theoretic Framework for Meta-Learning Generalization
by: Wen, Wen, et al.
Published: (2025)
by: Wen, Wen, et al.
Published: (2025)
Refined Risk Bounds for Unbounded Losses via Transductive Priors
by: Qian, Jian, et al.
Published: (2024)
by: Qian, Jian, et al.
Published: (2024)
Switching the Loss Reduces the Cost in Batch (Offline) Reinforcement Learning
by: Ayoub, Alex, et al.
Published: (2024)
by: Ayoub, Alex, et al.
Published: (2024)
The Curious Price of Distributional Robustness in Reinforcement Learning with a Generative Model
by: Shi, Laixi, et al.
Published: (2023)
by: Shi, Laixi, et al.
Published: (2023)
A Central Limit Theorem for the permutation importance measure
by: Föge, Nico, et al.
Published: (2024)
by: Föge, Nico, et al.
Published: (2024)
Functional Linear Regression of Cumulative Distribution Functions
by: Zhang, Qian, et al.
Published: (2022)
by: Zhang, Qian, et al.
Published: (2022)
MetaCURL: Non-stationary Concave Utility Reinforcement Learning
by: Moreno, Bianca Marin, et al.
Published: (2024)
by: Moreno, Bianca Marin, et al.
Published: (2024)
When Are Trade-Off Functions Testable from Finite Samples?
by: Shi, Kaining, et al.
Published: (2026)
by: Shi, Kaining, et al.
Published: (2026)
Conformal Risk Control for Non-Monotonic Losses
by: Angelopoulos, Anastasios N.
Published: (2026)
by: Angelopoulos, Anastasios N.
Published: (2026)
MoMA: Model-based Mirror Ascent for Offline Reinforcement Learning
by: Hong, Mao, et al.
Published: (2024)
by: Hong, Mao, et al.
Published: (2024)
Nonparametric Jackknife Instrumental Variable Estimation and Confounding Robust Surrogate Indices
by: Bibaut, Aurélien, et al.
Published: (2024)
by: Bibaut, Aurélien, et al.
Published: (2024)
Adaptive Refinement Protocols for Distributed Distribution Estimation under $\ell^p$-Losses
by: Yuan, Deheng, et al.
Published: (2024)
by: Yuan, Deheng, et al.
Published: (2024)
Physics-informed kernel learning
by: Doumèche, Nathan, et al.
Published: (2024)
by: Doumèche, Nathan, et al.
Published: (2024)
Finite-Dimensional Gaussian Approximation for Deep Neural Networks: Universality in Random Weights
by: Balasubramanian, Krishnakumar, et al.
Published: (2025)
by: Balasubramanian, Krishnakumar, et al.
Published: (2025)
On Uncertainty Calibration for Equivariant Functions
by: Berman, Edward, et al.
Published: (2025)
by: Berman, Edward, et al.
Published: (2025)
Similar Items
-
Efficient Inference for Inverse Reinforcement Learning and Dynamic Discrete Choice Models
by: van der Laan, Lars, et al.
Published: (2025) -
Automatic Debiased Machine Learning for Smooth Functionals of Nonparametric M-Estimands
by: van der Laan, Lars, et al.
Published: (2025) -
Demistifying Inference after Adaptive Experiments
by: Bibaut, Aurélien, et al.
Published: (2024) -
Exploration in the Limit
by: Cho, Brian M., et al.
Published: (2025) -
Clustered Switchback Designs for Experimentation Under Spatio-temporal Interference
by: Jia, Su, et al.
Published: (2023)