Adaptive Discretization in Online Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Sinclair, Sean R., Banerjee, Siddhartha, Yu, Christina Lee |
|---|---|
| Format: | Preprint |
| Published: |
2021
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reinforcement Learning in MDPs with Information-Ordered Policies
by: Zhang, Zhongjun, et al.
Published: (2025)
by: Zhang, Zhongjun, et al.
Published: (2025)
Ambiguous Online Learning
by: Kosoy, Vanessa
Published: (2025)
by: Kosoy, Vanessa
Published: (2025)
Regret Bounds for Robust Online Decision Making
by: Appel, Alexander, et al.
Published: (2025)
by: Appel, Alexander, et al.
Published: (2025)
Aligning Inductive Bias for Data-Efficient Generalization in State Space Models
by: Chen, Qiyu, et al.
Published: (2025)
by: Chen, Qiyu, et al.
Published: (2025)
Agnostic Learning under Targeted Poisoning: Optimal Rates and the Role of Randomness
by: Chornomaz, Bogdan, et al.
Published: (2025)
by: Chornomaz, Bogdan, et al.
Published: (2025)
Backpropagation Through Time For Networks With Long-Term Dependencies
by: Bird, George, et al.
Published: (2021)
by: Bird, George, et al.
Published: (2021)
Frequency Principle: Fourier Analysis Sheds Light on Deep Neural Networks
by: Xu, Zhi-Qin John, et al.
Published: (2019)
by: Xu, Zhi-Qin John, et al.
Published: (2019)
Reinforcement Learning for Reachability: Guaranteeing Asymptotic Optimality
by: Palasamudram, Amogh, et al.
Published: (2026)
by: Palasamudram, Amogh, et al.
Published: (2026)
Retrieval-Augmented Memory for Online Learning
by: Du, Wenzhang
Published: (2025)
by: Du, Wenzhang
Published: (2025)
Sparse Knowledge Distillation: A Mathematical Framework for Probability-Domain Temperature Scaling and Multi-Stage Compression
by: Flouro, Aaron R., et al.
Published: (2026)
by: Flouro, Aaron R., et al.
Published: (2026)
Margin in Abstract Spaces
by: Ashlagi, Yair, et al.
Published: (2026)
by: Ashlagi, Yair, et al.
Published: (2026)
Statistical Guarantees for Lifelong Reinforcement Learning using PAC-Bayes Theory
by: Zhang, Zhi, et al.
Published: (2024)
by: Zhang, Zhi, et al.
Published: (2024)
Golden Handcuffs make safer AI agents
by: Ebtekar, Aram, et al.
Published: (2026)
by: Ebtekar, Aram, et al.
Published: (2026)
Mitigating Catastrophic Forgetting in Streaming Generative and Predictive Learning via Stateful Replay
by: Du, Wenzhang
Published: (2025)
by: Du, Wenzhang
Published: (2025)
Superior Scoring Rules for Probabilistic Evaluation of Single-Label Multi-Class Classification Tasks
by: Ahmadian, Rouhollah, et al.
Published: (2024)
by: Ahmadian, Rouhollah, et al.
Published: (2024)
Tricks and Plug-ins for Gradient Boosting with Transformers
by: Fang, Biyi, et al.
Published: (2025)
by: Fang, Biyi, et al.
Published: (2025)
The Current and Future Perspectives of Zinc Oxide Nanoparticles in the Treatment of Diabetes Mellitus
by: Yousaf, Iqra
Published: (2024)
by: Yousaf, Iqra
Published: (2024)
MMD-Balls as Credal Sets: A PAC-Bayesian Framework for Epistemic Uncertainty in Test-Time Adaptation
by: Ariq, Ahanaf Hasan
Published: (2026)
by: Ariq, Ahanaf Hasan
Published: (2026)
Inductive Venn-Abers and related regressors
by: Petej, Ivan, et al.
Published: (2026)
by: Petej, Ivan, et al.
Published: (2026)
Aggregation in conformal e-classification
by: Vovk, Vladimir
Published: (2026)
by: Vovk, Vladimir
Published: (2026)
Theoretical Analysis of Positional Encodings in Transformer Models: Impact on Expressiveness and Generalization
by: Li, Yin
Published: (2025)
by: Li, Yin
Published: (2025)
Optimistic Feasible Search for Closed-Loop Fair Threshold Decision-Making
by: Du, Wenzhang
Published: (2025)
by: Du, Wenzhang
Published: (2025)
Online Decision Making with Generative Action Sets
by: Xu, Jianyu, et al.
Published: (2025)
by: Xu, Jianyu, et al.
Published: (2025)
Advances in Set Function Learning: A Survey of Techniques and Applications
by: Xie, Jiahao, et al.
Published: (2025)
by: Xie, Jiahao, et al.
Published: (2025)
Position: Uncertainty Quantification in LLMs is Just Unsupervised Clustering
by: Chen, Tiejin, et al.
Published: (2026)
by: Chen, Tiejin, et al.
Published: (2026)
Energy-Efficient Information Representation in MNIST Classification Using Biologically Inspired Learning
by: Stricker, Patrick, et al.
Published: (2026)
by: Stricker, Patrick, et al.
Published: (2026)
ZetA: A Riemann Zeta-Scaled Extension of Adam for Deep Learning
by: BC, Samiksha
Published: (2025)
by: BC, Samiksha
Published: (2025)
Black-Box Uniform Stability for Non-Euclidean Empirical Risk Minimization
by: Vary, Simon, et al.
Published: (2024)
by: Vary, Simon, et al.
Published: (2024)
The two clocks and the innovation window: When and how generative models learn rules
by: Wang, Binxu, et al.
Published: (2026)
by: Wang, Binxu, et al.
Published: (2026)
Causal Direction from Convergence Time: Faster Training in the True Causal Direction
by: Tamim, Abdulrahman
Published: (2026)
by: Tamim, Abdulrahman
Published: (2026)
A Reduction from Delayed to Immediate Feedback for Online Convex Optimization with Improved Guarantees
by: Ryabchenko, Alexander, et al.
Published: (2026)
by: Ryabchenko, Alexander, et al.
Published: (2026)
Hallucinations Live in Variance
by: Flouro, Aaron R., et al.
Published: (2026)
by: Flouro, Aaron R., et al.
Published: (2026)
Diagnosing Multi-step Reasoning Failures in Black-box LLMs via Stepwise Confidence Attribution
by: Liu, Xiaoou, et al.
Published: (2026)
by: Liu, Xiaoou, et al.
Published: (2026)
Modularity in Transformers: Investigating Neuron Separability & Specialization
by: Pochinkov, Nicholas, et al.
Published: (2024)
by: Pochinkov, Nicholas, et al.
Published: (2024)
On the Equivalence of Regression and Classification
by: Jayadeva, et al.
Published: (2025)
by: Jayadeva, et al.
Published: (2025)
InhibiDistilbert: Knowledge Distillation for a ReLU and Addition-based Transformer
by: Zhang, Tony, et al.
Published: (2025)
by: Zhang, Tony, et al.
Published: (2025)
Binarized Neural Networks Converge Toward Algorithmic Simplicity: Empirical Support for the Learning-as-Compression Hypothesis
by: Sakabe, Eduardo Y., et al.
Published: (2025)
by: Sakabe, Eduardo Y., et al.
Published: (2025)
Pair Correlation Factor and the Sample Complexity of Gaussian Mixtures
by: Aryan, Farzad
Published: (2025)
by: Aryan, Farzad
Published: (2025)
Understanding the Nature of Generative AI as Threshold Logic in High-Dimensional Space
by: Levin, Ilya
Published: (2026)
by: Levin, Ilya
Published: (2026)
Closed-Form Beta Distribution Estimation from Sparse Statistics with Random Forest Implicit Regularization
by: Landers, Jonathan R.
Published: (2025)
by: Landers, Jonathan R.
Published: (2025)
Similar Items
-
Reinforcement Learning in MDPs with Information-Ordered Policies
by: Zhang, Zhongjun, et al.
Published: (2025) -
Ambiguous Online Learning
by: Kosoy, Vanessa
Published: (2025) -
Regret Bounds for Robust Online Decision Making
by: Appel, Alexander, et al.
Published: (2025) -
Aligning Inductive Bias for Data-Efficient Generalization in State Space Models
by: Chen, Qiyu, et al.
Published: (2025) -
Agnostic Learning under Targeted Poisoning: Optimal Rates and the Role of Randomness
by: Chornomaz, Bogdan, et al.
Published: (2025)