Sparse Optimistic Information Directed Sampling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Schwartz, Ludovic, Flynn, Hamish, Neu, Gergely |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Optimistic Information Directed Sampling
von: Neu, Gergely, et al.
Veröffentlicht: (2024)
von: Neu, Gergely, et al.
Veröffentlicht: (2024)
Linear Bandits with Non-i.i.d. Noise
von: Abélès, Baptiste, et al.
Veröffentlicht: (2025)
von: Abélès, Baptiste, et al.
Veröffentlicht: (2025)
Optimistically Optimistic Exploration for Provably Efficient Infinite-Horizon Reinforcement and Imitation Learning
von: Moulin, Antoine, et al.
Veröffentlicht: (2025)
von: Moulin, Antoine, et al.
Veröffentlicht: (2025)
Confidence Sequences for Generalized Linear Models via Regret Analysis
von: Clerico, Eugenio, et al.
Veröffentlicht: (2025)
von: Clerico, Eugenio, et al.
Veröffentlicht: (2025)
Relative Information Gain and Gaussian Process Regression
von: Flynn, Hamish
Veröffentlicht: (2025)
von: Flynn, Hamish
Veröffentlicht: (2025)
Distances for Markov chains from sample streams
von: Calo, Sergio, et al.
Veröffentlicht: (2025)
von: Calo, Sergio, et al.
Veröffentlicht: (2025)
Bisimulation Metrics are Optimal Transport Distances, and Can be Computed Efficiently
von: Calo, Sergio, et al.
Veröffentlicht: (2024)
von: Calo, Sergio, et al.
Veröffentlicht: (2024)
Sparse Nonparametric Contextual Bandits
von: Flynn, Hamish, et al.
Veröffentlicht: (2025)
von: Flynn, Hamish, et al.
Veröffentlicht: (2025)
Tighter Confidence Bounds for Sequential Kernel Regression
von: Flynn, Hamish, et al.
Veröffentlicht: (2024)
von: Flynn, Hamish, et al.
Veröffentlicht: (2024)
Online combinatorial optimization with stochastic decision sets and adversarial losses
von: Neu, Gergely, et al.
Veröffentlicht: (2026)
von: Neu, Gergely, et al.
Veröffentlicht: (2026)
Offline RL via Feature-Occupancy Gradient Ascent
von: Neu, Gergely, et al.
Veröffentlicht: (2024)
von: Neu, Gergely, et al.
Veröffentlicht: (2024)
Online-to-PAC Conversions: Generalization Bounds via Regret Analysis
von: Lugosi, Gábor, et al.
Veröffentlicht: (2023)
von: Lugosi, Gábor, et al.
Veröffentlicht: (2023)
Posterior Sampling Reinforcement Learning with Gaussian Processes for Continuous Control: Sublinear Regret Bounds for Unbounded State Spaces
von: Flynn, Hamish, et al.
Veröffentlicht: (2026)
von: Flynn, Hamish, et al.
Veröffentlicht: (2026)
Dealing with unbounded gradients in stochastic saddle-point optimization
von: Neu, Gergely, et al.
Veröffentlicht: (2024)
von: Neu, Gergely, et al.
Veröffentlicht: (2024)
Inverse Q-Learning Done Right: Offline Imitation Learning in $Q^π$-Realizable MDPs
von: Moulin, Antoine, et al.
Veröffentlicht: (2025)
von: Moulin, Antoine, et al.
Veröffentlicht: (2025)
Online learning with Erdős-Rényi side-observation graphs
von: Kocák, Tomáš, et al.
Veröffentlicht: (2026)
von: Kocák, Tomáš, et al.
Veröffentlicht: (2026)
Online-to-PAC generalization bounds under graph-mixing dependencies
von: Abélès, Baptiste, et al.
Veröffentlicht: (2024)
von: Abélès, Baptiste, et al.
Veröffentlicht: (2024)
Online learning with noisy side observations
von: Kocák, Tomáš, et al.
Veröffentlicht: (2026)
von: Kocák, Tomáš, et al.
Veröffentlicht: (2026)
Generalization bounds for mixing processes via delayed online-to-PAC conversions
von: Abeles, Baptiste, et al.
Veröffentlicht: (2024)
von: Abeles, Baptiste, et al.
Veröffentlicht: (2024)
Improved Algorithms for Stochastic Linear Bandits Using Tail Bounds for Martingale Mixtures
von: Flynn, Hamish, et al.
Veröffentlicht: (2023)
von: Flynn, Hamish, et al.
Veröffentlicht: (2023)
Efficient learning by implicit exploration in bandit problems with side observations
von: Kocak, Tomas, et al.
Veröffentlicht: (2026)
von: Kocak, Tomas, et al.
Veröffentlicht: (2026)
Efficient Model-Based Reinforcement Learning Through Optimistic Thompson Sampling
von: Bayrooti, Jasmine, et al.
Veröffentlicht: (2024)
von: Bayrooti, Jasmine, et al.
Veröffentlicht: (2024)
COMPASS: COntinual Multilingual PEFT with Adaptive Semantic Sampling
von: Flynn, Noah
Veröffentlicht: (2026)
von: Flynn, Noah
Veröffentlicht: (2026)
Optimistic Thompson Sampling for No-Regret Learning in Unknown Games
von: Li, Yingru, et al.
Veröffentlicht: (2024)
von: Li, Yingru, et al.
Veröffentlicht: (2024)
Optimistic Policy Regularization
von: Pham, Mai, et al.
Veröffentlicht: (2026)
von: Pham, Mai, et al.
Veröffentlicht: (2026)
Sparse random hypergraphs: Non-backtracking spectra and community detection
von: Stephan, Ludovic, et al.
Veröffentlicht: (2022)
von: Stephan, Ludovic, et al.
Veröffentlicht: (2022)
Omega: Optimistic EMA Gradients
von: Ramirez, Juan, et al.
Veröffentlicht: (2023)
von: Ramirez, Juan, et al.
Veröffentlicht: (2023)
Optimistic critics can empower small actors
von: Mastikhina, Olya, et al.
Veröffentlicht: (2025)
von: Mastikhina, Olya, et al.
Veröffentlicht: (2025)
SOMBRL: Scalable and Optimistic Model-Based RL
von: Sukhija, Bhavya, et al.
Veröffentlicht: (2025)
von: Sukhija, Bhavya, et al.
Veröffentlicht: (2025)
Optimistic Task Inference for Behavior Foundation Models
von: Rupf, Thomas, et al.
Veröffentlicht: (2025)
von: Rupf, Thomas, et al.
Veröffentlicht: (2025)
Bayesian Optimistic Optimisation with Exponentially Decaying Regret
von: Tran-The, Hung, et al.
Veröffentlicht: (2021)
von: Tran-The, Hung, et al.
Veröffentlicht: (2021)
Optimistic Dual Averaging Unifies Modern Optimizers
von: Pethick, Thomas, et al.
Veröffentlicht: (2026)
von: Pethick, Thomas, et al.
Veröffentlicht: (2026)
Optimistic Learning for Communication Networks
von: Iosifidis, George, et al.
Veröffentlicht: (2025)
von: Iosifidis, George, et al.
Veröffentlicht: (2025)
Optimistic Model Rollouts for Pessimistic Offline Policy Optimization
von: Zhai, Yuanzhao, et al.
Veröffentlicht: (2024)
von: Zhai, Yuanzhao, et al.
Veröffentlicht: (2024)
Optimistic Reinforcement Learning with Quantile Objectives
von: Alipour-Vaezi, Mohammad, et al.
Veröffentlicht: (2025)
von: Alipour-Vaezi, Mohammad, et al.
Veröffentlicht: (2025)
Optimistic Multi-Agent Policy Gradient
von: Zhao, Wenshuai, et al.
Veröffentlicht: (2023)
von: Zhao, Wenshuai, et al.
Veröffentlicht: (2023)
On Stability in Optimistic Bilevel Optimization
von: Royset, Johannes O.
Veröffentlicht: (2024)
von: Royset, Johannes O.
Veröffentlicht: (2024)
Optimistic Interior Point Methods for Sequential Hypothesis Testing by Betting
von: Chen, Can, et al.
Veröffentlicht: (2025)
von: Chen, Can, et al.
Veröffentlicht: (2025)
Optimistic Q-learning for average reward and episodic reinforcement learning
von: Agrawal, Priyank, et al.
Veröffentlicht: (2024)
von: Agrawal, Priyank, et al.
Veröffentlicht: (2024)
Optimistic Policy Optimization is Provably Efficient in Non-stationary MDPs
von: Zhong, Han, et al.
Veröffentlicht: (2021)
von: Zhong, Han, et al.
Veröffentlicht: (2021)
Ähnliche Einträge
-
Optimistic Information Directed Sampling
von: Neu, Gergely, et al.
Veröffentlicht: (2024) -
Linear Bandits with Non-i.i.d. Noise
von: Abélès, Baptiste, et al.
Veröffentlicht: (2025) -
Optimistically Optimistic Exploration for Provably Efficient Infinite-Horizon Reinforcement and Imitation Learning
von: Moulin, Antoine, et al.
Veröffentlicht: (2025) -
Confidence Sequences for Generalized Linear Models via Regret Analysis
von: Clerico, Eugenio, et al.
Veröffentlicht: (2025) -
Relative Information Gain and Gaussian Process Regression
von: Flynn, Hamish
Veröffentlicht: (2025)