Reinforcement Learning with Action-Triggered Observations
Fuente:
arXiv
Guardado en:
| Autores principales: | Ryabchenko, Alexander, Mou, Wenlong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Extensions of the regret-minimization algorithm for optimal design
por: Chen, Youguang, et al.
Publicado: (2025)
por: Chen, Youguang, et al.
Publicado: (2025)
Credal Bayesian Deep Learning
por: Caprio, Michele, et al.
Publicado: (2023)
por: Caprio, Michele, et al.
Publicado: (2023)
On Lai's Upper Confidence Bound in Multi-Armed Bandits
por: Ren, Huachen, et al.
Publicado: (2024)
por: Ren, Huachen, et al.
Publicado: (2024)
Distributed Parallel Structure-Aware Presolving for Arrowhead Linear Programs
por: Kempke, Nils-Christian, et al.
Publicado: (2026)
por: Kempke, Nils-Christian, et al.
Publicado: (2026)
Learning-Augmented Algorithms for MTS with Bandit Access to Multiple Predictors
por: Coşa, Matei Gabriel, et al.
Publicado: (2025)
por: Coşa, Matei Gabriel, et al.
Publicado: (2025)
FedSTaS: Client Stratification and Client Level Sampling for Efficient Federated Learning
por: Slessor, Jordan, et al.
Publicado: (2024)
por: Slessor, Jordan, et al.
Publicado: (2024)
Validity and efficiency of the conformal CUSUM procedure
por: Vovk, Vladimir, et al.
Publicado: (2024)
por: Vovk, Vladimir, et al.
Publicado: (2024)
An Analysis of Optimizer Choice on Energy Efficiency and Performance in Neural Network Training
por: Almog, Tom
Publicado: (2025)
por: Almog, Tom
Publicado: (2025)
Credal and Interval Deep Evidential Classifications
por: Caprio, Michele, et al.
Publicado: (2025)
por: Caprio, Michele, et al.
Publicado: (2025)
Evolutionary Computation as Natural Generative AI
por: Shi, Yaxin, et al.
Publicado: (2025)
por: Shi, Yaxin, et al.
Publicado: (2025)
Improved Regret Guarantees for Online Mirror Descent using a Portfolio of Mirror Maps
por: Gupta, Swati, et al.
Publicado: (2026)
por: Gupta, Swati, et al.
Publicado: (2026)
MSTN: A Lightweight and Fast Model for General TimeSeries Analysis
por: Shevtekar, Sumit S, et al.
Publicado: (2025)
por: Shevtekar, Sumit S, et al.
Publicado: (2025)
Time Series Analysis by State Space Learning
por: Ramos, André, et al.
Publicado: (2024)
por: Ramos, André, et al.
Publicado: (2024)
Decorrelation, Diversity, and Emergent Intelligence: The Isomorphism Between Social Insect Colonies and Ensemble Machine Learning
por: Fokoué, Ernest, et al.
Publicado: (2026)
por: Fokoué, Ernest, et al.
Publicado: (2026)
On the Optimality of the Oja's Algorithm for Online PCA
por: Liang, Xin
Publicado: (2021)
por: Liang, Xin
Publicado: (2021)
Conformal e-prediction
por: Vovk, Vladimir
Publicado: (2020)
por: Vovk, Vladimir
Publicado: (2020)
Differentially Private Fisher Randomization Tests for Binary Outcomes
por: Sun, Qingyang, et al.
Publicado: (2025)
por: Sun, Qingyang, et al.
Publicado: (2025)
On the boundedness of the sequence generated by minibatch stochastic gradient descent
por: Bauschke, Heinz H., et al.
Publicado: (2025)
por: Bauschke, Heinz H., et al.
Publicado: (2025)
A Function-Space Stability Boundary for Generalization in Interpolating Learning Systems
por: Katende, Ronald
Publicado: (2026)
por: Katende, Ronald
Publicado: (2026)
Enhancing Diversity in Multi-objective Feature Selection
por: Miyandoab, Sevil Zanjani, et al.
Publicado: (2024)
por: Miyandoab, Sevil Zanjani, et al.
Publicado: (2024)
Approximating k-Center via Farthest-First on $δ$-Covers
por: Wilson, Jason R.
Publicado: (2026)
por: Wilson, Jason R.
Publicado: (2026)
Causality as the Statistical Conscience of Artificial Intelligence: From Pearl's Ladder to Trustworthy Machines
por: Fokoué, Ernest
Publicado: (2026)
por: Fokoué, Ernest
Publicado: (2026)
Conformal e-testing
por: Vovk, Vladimir, et al.
Publicado: (2020)
por: Vovk, Vladimir, et al.
Publicado: (2020)
On the existence of minimizers in shallow residual ReLU neural network optimization landscapes
por: Dereich, Steffen, et al.
Publicado: (2023)
por: Dereich, Steffen, et al.
Publicado: (2023)
On the existence of optimal shallow feedforward networks with ReLU activation
por: Dereich, Steffen, et al.
Publicado: (2023)
por: Dereich, Steffen, et al.
Publicado: (2023)
On the on-line coloring of unit interval graphs with proper interval representation
por: Curbelo, Israel R., et al.
Publicado: (2024)
por: Curbelo, Israel R., et al.
Publicado: (2024)
Symbolic Branch Networks: Tree-Inherited Neural Models for Interpretable Multiclass Classification
por: Rodríguez-Salas, Dalia
Publicado: (2025)
por: Rodríguez-Salas, Dalia
Publicado: (2025)
Inductive randomness predictors: beyond conformal
por: Vovk, Vladimir
Publicado: (2025)
por: Vovk, Vladimir
Publicado: (2025)
A Reduction from Delayed to Immediate Feedback for Online Convex Optimization with Improved Guarantees
por: Ryabchenko, Alexander, et al.
Publicado: (2026)
por: Ryabchenko, Alexander, et al.
Publicado: (2026)
Online Trading as a Secretary Problem Variant
por: Chen, Xujin, et al.
Publicado: (2026)
por: Chen, Xujin, et al.
Publicado: (2026)
Dirichlet Scale Mixture Priors for Bayesian Neural Networks
por: Arnstad, August, et al.
Publicado: (2026)
por: Arnstad, August, et al.
Publicado: (2026)
A Framework for Scalable Heterogeneous Multi-Agent Adversarial Reinforcement Learning in IsaacLab
por: Peterson, Isaac, et al.
Publicado: (2025)
por: Peterson, Isaac, et al.
Publicado: (2025)
Universality of conformal prediction under the assumption of randomness
por: Vovk, Vladimir
Publicado: (2025)
por: Vovk, Vladimir
Publicado: (2025)
Conformal e-prediction in the presence of confounding
por: Vovk, Vladimir, et al.
Publicado: (2026)
por: Vovk, Vladimir, et al.
Publicado: (2026)
Optimistic Training and Convergence of Q-Learning -- Extended Version
por: Mehta, Prashant, et al.
Publicado: (2026)
por: Mehta, Prashant, et al.
Publicado: (2026)
Stability and Sensitivity Analysis of Relative Temporal-Difference Learning: Extended Version
por: Sakha, Masoud S., et al.
Publicado: (2026)
por: Sakha, Masoud S., et al.
Publicado: (2026)
Power Homotopy for Zeroth-Order Non-Convex Optimizations
por: Xu, Chen
Publicado: (2025)
por: Xu, Chen
Publicado: (2025)
Global Optimization with A Power-Transformed Objective and Gaussian Smoothing
por: Xu, Chen
Publicado: (2024)
por: Xu, Chen
Publicado: (2024)
A genetic algorithm to search the space of Ehrhart $h^*$-vectors
por: Balletti, Gabriele
Publicado: (2023)
por: Balletti, Gabriele
Publicado: (2023)
Computing conservative probabilities of rare events with surrogates
por: Bousquet, Nicolas
Publicado: (2024)
por: Bousquet, Nicolas
Publicado: (2024)
Ejemplares similares
-
Extensions of the regret-minimization algorithm for optimal design
por: Chen, Youguang, et al.
Publicado: (2025) -
Credal Bayesian Deep Learning
por: Caprio, Michele, et al.
Publicado: (2023) -
On Lai's Upper Confidence Bound in Multi-Armed Bandits
por: Ren, Huachen, et al.
Publicado: (2024) -
Distributed Parallel Structure-Aware Presolving for Arrowhead Linear Programs
por: Kempke, Nils-Christian, et al.
Publicado: (2026) -
Learning-Augmented Algorithms for MTS with Bandit Access to Multiple Predictors
por: Coşa, Matei Gabriel, et al.
Publicado: (2025)