Reinforcement Learning with Action-Triggered Observations
Fuente:
arXiv
Salvato in:
| Autori principali: | Ryabchenko, Alexander, Mou, Wenlong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Extensions of the regret-minimization algorithm for optimal design
di: Chen, Youguang, et al.
Pubblicazione: (2025)
di: Chen, Youguang, et al.
Pubblicazione: (2025)
Credal Bayesian Deep Learning
di: Caprio, Michele, et al.
Pubblicazione: (2023)
di: Caprio, Michele, et al.
Pubblicazione: (2023)
On Lai's Upper Confidence Bound in Multi-Armed Bandits
di: Ren, Huachen, et al.
Pubblicazione: (2024)
di: Ren, Huachen, et al.
Pubblicazione: (2024)
Distributed Parallel Structure-Aware Presolving for Arrowhead Linear Programs
di: Kempke, Nils-Christian, et al.
Pubblicazione: (2026)
di: Kempke, Nils-Christian, et al.
Pubblicazione: (2026)
Learning-Augmented Algorithms for MTS with Bandit Access to Multiple Predictors
di: Coşa, Matei Gabriel, et al.
Pubblicazione: (2025)
di: Coşa, Matei Gabriel, et al.
Pubblicazione: (2025)
FedSTaS: Client Stratification and Client Level Sampling for Efficient Federated Learning
di: Slessor, Jordan, et al.
Pubblicazione: (2024)
di: Slessor, Jordan, et al.
Pubblicazione: (2024)
Validity and efficiency of the conformal CUSUM procedure
di: Vovk, Vladimir, et al.
Pubblicazione: (2024)
di: Vovk, Vladimir, et al.
Pubblicazione: (2024)
An Analysis of Optimizer Choice on Energy Efficiency and Performance in Neural Network Training
di: Almog, Tom
Pubblicazione: (2025)
di: Almog, Tom
Pubblicazione: (2025)
Credal and Interval Deep Evidential Classifications
di: Caprio, Michele, et al.
Pubblicazione: (2025)
di: Caprio, Michele, et al.
Pubblicazione: (2025)
Evolutionary Computation as Natural Generative AI
di: Shi, Yaxin, et al.
Pubblicazione: (2025)
di: Shi, Yaxin, et al.
Pubblicazione: (2025)
Improved Regret Guarantees for Online Mirror Descent using a Portfolio of Mirror Maps
di: Gupta, Swati, et al.
Pubblicazione: (2026)
di: Gupta, Swati, et al.
Pubblicazione: (2026)
MSTN: A Lightweight and Fast Model for General TimeSeries Analysis
di: Shevtekar, Sumit S, et al.
Pubblicazione: (2025)
di: Shevtekar, Sumit S, et al.
Pubblicazione: (2025)
Time Series Analysis by State Space Learning
di: Ramos, André, et al.
Pubblicazione: (2024)
di: Ramos, André, et al.
Pubblicazione: (2024)
Decorrelation, Diversity, and Emergent Intelligence: The Isomorphism Between Social Insect Colonies and Ensemble Machine Learning
di: Fokoué, Ernest, et al.
Pubblicazione: (2026)
di: Fokoué, Ernest, et al.
Pubblicazione: (2026)
On the Optimality of the Oja's Algorithm for Online PCA
di: Liang, Xin
Pubblicazione: (2021)
di: Liang, Xin
Pubblicazione: (2021)
Conformal e-prediction
di: Vovk, Vladimir
Pubblicazione: (2020)
di: Vovk, Vladimir
Pubblicazione: (2020)
Differentially Private Fisher Randomization Tests for Binary Outcomes
di: Sun, Qingyang, et al.
Pubblicazione: (2025)
di: Sun, Qingyang, et al.
Pubblicazione: (2025)
On the boundedness of the sequence generated by minibatch stochastic gradient descent
di: Bauschke, Heinz H., et al.
Pubblicazione: (2025)
di: Bauschke, Heinz H., et al.
Pubblicazione: (2025)
A Function-Space Stability Boundary for Generalization in Interpolating Learning Systems
di: Katende, Ronald
Pubblicazione: (2026)
di: Katende, Ronald
Pubblicazione: (2026)
Enhancing Diversity in Multi-objective Feature Selection
di: Miyandoab, Sevil Zanjani, et al.
Pubblicazione: (2024)
di: Miyandoab, Sevil Zanjani, et al.
Pubblicazione: (2024)
Approximating k-Center via Farthest-First on $δ$-Covers
di: Wilson, Jason R.
Pubblicazione: (2026)
di: Wilson, Jason R.
Pubblicazione: (2026)
Causality as the Statistical Conscience of Artificial Intelligence: From Pearl's Ladder to Trustworthy Machines
di: Fokoué, Ernest
Pubblicazione: (2026)
di: Fokoué, Ernest
Pubblicazione: (2026)
Conformal e-testing
di: Vovk, Vladimir, et al.
Pubblicazione: (2020)
di: Vovk, Vladimir, et al.
Pubblicazione: (2020)
On the existence of minimizers in shallow residual ReLU neural network optimization landscapes
di: Dereich, Steffen, et al.
Pubblicazione: (2023)
di: Dereich, Steffen, et al.
Pubblicazione: (2023)
On the existence of optimal shallow feedforward networks with ReLU activation
di: Dereich, Steffen, et al.
Pubblicazione: (2023)
di: Dereich, Steffen, et al.
Pubblicazione: (2023)
On the on-line coloring of unit interval graphs with proper interval representation
di: Curbelo, Israel R., et al.
Pubblicazione: (2024)
di: Curbelo, Israel R., et al.
Pubblicazione: (2024)
Symbolic Branch Networks: Tree-Inherited Neural Models for Interpretable Multiclass Classification
di: Rodríguez-Salas, Dalia
Pubblicazione: (2025)
di: Rodríguez-Salas, Dalia
Pubblicazione: (2025)
Inductive randomness predictors: beyond conformal
di: Vovk, Vladimir
Pubblicazione: (2025)
di: Vovk, Vladimir
Pubblicazione: (2025)
A Reduction from Delayed to Immediate Feedback for Online Convex Optimization with Improved Guarantees
di: Ryabchenko, Alexander, et al.
Pubblicazione: (2026)
di: Ryabchenko, Alexander, et al.
Pubblicazione: (2026)
Online Trading as a Secretary Problem Variant
di: Chen, Xujin, et al.
Pubblicazione: (2026)
di: Chen, Xujin, et al.
Pubblicazione: (2026)
Dirichlet Scale Mixture Priors for Bayesian Neural Networks
di: Arnstad, August, et al.
Pubblicazione: (2026)
di: Arnstad, August, et al.
Pubblicazione: (2026)
A Framework for Scalable Heterogeneous Multi-Agent Adversarial Reinforcement Learning in IsaacLab
di: Peterson, Isaac, et al.
Pubblicazione: (2025)
di: Peterson, Isaac, et al.
Pubblicazione: (2025)
Universality of conformal prediction under the assumption of randomness
di: Vovk, Vladimir
Pubblicazione: (2025)
di: Vovk, Vladimir
Pubblicazione: (2025)
Conformal e-prediction in the presence of confounding
di: Vovk, Vladimir, et al.
Pubblicazione: (2026)
di: Vovk, Vladimir, et al.
Pubblicazione: (2026)
Optimistic Training and Convergence of Q-Learning -- Extended Version
di: Mehta, Prashant, et al.
Pubblicazione: (2026)
di: Mehta, Prashant, et al.
Pubblicazione: (2026)
Stability and Sensitivity Analysis of Relative Temporal-Difference Learning: Extended Version
di: Sakha, Masoud S., et al.
Pubblicazione: (2026)
di: Sakha, Masoud S., et al.
Pubblicazione: (2026)
Power Homotopy for Zeroth-Order Non-Convex Optimizations
di: Xu, Chen
Pubblicazione: (2025)
di: Xu, Chen
Pubblicazione: (2025)
Global Optimization with A Power-Transformed Objective and Gaussian Smoothing
di: Xu, Chen
Pubblicazione: (2024)
di: Xu, Chen
Pubblicazione: (2024)
A genetic algorithm to search the space of Ehrhart $h^*$-vectors
di: Balletti, Gabriele
Pubblicazione: (2023)
di: Balletti, Gabriele
Pubblicazione: (2023)
Computing conservative probabilities of rare events with surrogates
di: Bousquet, Nicolas
Pubblicazione: (2024)
di: Bousquet, Nicolas
Pubblicazione: (2024)
Documenti analoghi
-
Extensions of the regret-minimization algorithm for optimal design
di: Chen, Youguang, et al.
Pubblicazione: (2025) -
Credal Bayesian Deep Learning
di: Caprio, Michele, et al.
Pubblicazione: (2023) -
On Lai's Upper Confidence Bound in Multi-Armed Bandits
di: Ren, Huachen, et al.
Pubblicazione: (2024) -
Distributed Parallel Structure-Aware Presolving for Arrowhead Linear Programs
di: Kempke, Nils-Christian, et al.
Pubblicazione: (2026) -
Learning-Augmented Algorithms for MTS with Bandit Access to Multiple Predictors
di: Coşa, Matei Gabriel, et al.
Pubblicazione: (2025)