Model selection for behavioral learning data and applications to contextual bandits
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Aubert, Julien, Köhler, Louis, Lehéricy, Luc, Mezzadri, Giulia, Reynaud-Bouret, Patricia |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
General oracle inequalities for a penalized log-likelihood criterion based on non-stationary data
von: Aubert, Julien, et al.
Veröffentlicht: (2024)
von: Aubert, Julien, et al.
Veröffentlicht: (2024)
Spiking Neural Models for Decision-Making Tasks with Learning
von: Jaffard, Sophie, et al.
Veröffentlicht: (2025)
von: Jaffard, Sophie, et al.
Veröffentlicht: (2025)
CHANI: Correlation-based Hawkes Aggregation of Neurons with bio-Inspiration
von: Jaffard, Sophie, et al.
Veröffentlicht: (2024)
von: Jaffard, Sophie, et al.
Veröffentlicht: (2024)
Optimal cross-learning for contextual bandits with unknown context distributions
von: Schneider, Jon, et al.
Veröffentlicht: (2024)
von: Schneider, Jon, et al.
Veröffentlicht: (2024)
Quantum contextual bandits and recommender systems for quantum data
von: Brahmachari, Shrigyan, et al.
Veröffentlicht: (2023)
von: Brahmachari, Shrigyan, et al.
Veröffentlicht: (2023)
VITS : Variational Inference Thompson Sampling for contextual bandits
von: Clavier, Pierre, et al.
Veröffentlicht: (2023)
von: Clavier, Pierre, et al.
Veröffentlicht: (2023)
Best-of-Both Worlds for linear contextual bandits with paid observations
von: Boyer, Nathan, et al.
Veröffentlicht: (2025)
von: Boyer, Nathan, et al.
Veröffentlicht: (2025)
Leveraging heterogeneous spillover in maximizing contextual bandit rewards
von: Faruk, Ahmed Sayeed, et al.
Veröffentlicht: (2023)
von: Faruk, Ahmed Sayeed, et al.
Veröffentlicht: (2023)
Anytime-valid off-policy inference for contextual bandits
von: Waudby-Smith, Ian, et al.
Veröffentlicht: (2022)
von: Waudby-Smith, Ian, et al.
Veröffentlicht: (2022)
A conversion theorem and minimax optimality for continuum contextual bandits
von: Akhavan, Arya, et al.
Veröffentlicht: (2024)
von: Akhavan, Arya, et al.
Veröffentlicht: (2024)
Vector preference-based contextual bandits under distributional shifts
von: Shukla, Apurv, et al.
Veröffentlicht: (2025)
von: Shukla, Apurv, et al.
Veröffentlicht: (2025)
Precision autotuning for linear solvers via contextual bandit-based RL
von: Carson, Erin, et al.
Veröffentlicht: (2026)
von: Carson, Erin, et al.
Veröffentlicht: (2026)
Online learning in bandits with predicted context
von: Guo, Yongyi, et al.
Veröffentlicht: (2023)
von: Guo, Yongyi, et al.
Veröffentlicht: (2023)
A single algorithm for both restless and rested rotting bandits
von: Seznec, Julien, et al.
Veröffentlicht: (2026)
von: Seznec, Julien, et al.
Veröffentlicht: (2026)
Stochastic contextual bandits with graph feedback: from independence number to MAS number
von: Wen, Yuxiao, et al.
Veröffentlicht: (2024)
von: Wen, Yuxiao, et al.
Veröffentlicht: (2024)
Spectral bandits for smooth graph functions with applications in recommender systems
von: Kocák, Tomáš, et al.
Veröffentlicht: (2026)
von: Kocák, Tomáš, et al.
Veröffentlicht: (2026)
Covariance-adapting algorithm for semi-bandits with application to sparse rewards
von: Perrault, Pierre, et al.
Veröffentlicht: (2026)
von: Perrault, Pierre, et al.
Veröffentlicht: (2026)
Extreme bandits
von: Carpentier, Alexandra, et al.
Veröffentlicht: (2026)
von: Carpentier, Alexandra, et al.
Veröffentlicht: (2026)
Efficient learning by implicit exploration in bandit problems with side observations
von: Kocak, Tomas, et al.
Veröffentlicht: (2026)
von: Kocak, Tomas, et al.
Veröffentlicht: (2026)
Reinforcement learning with combinatorial actions for coupled restless bandits
von: Xu, Lily, et al.
Veröffentlicht: (2025)
von: Xu, Lily, et al.
Veröffentlicht: (2025)
Spectral bandits
von: Kocák, Tomáš, et al.
Veröffentlicht: (2026)
von: Kocák, Tomáš, et al.
Veröffentlicht: (2026)
Active clustering with bandit feedback
von: Thuot, Victor, et al.
Veröffentlicht: (2024)
von: Thuot, Victor, et al.
Veröffentlicht: (2024)
Neural Coding as a Statistical Testing Problem
von: Ost, Guilherme, et al.
Veröffentlicht: (2022)
von: Ost, Guilherme, et al.
Veröffentlicht: (2022)
Pair-Matching: Links Prediction with Adaptive Queries
von: Giraud, Christophe, et al.
Veröffentlicht: (2019)
von: Giraud, Christophe, et al.
Veröffentlicht: (2019)
Small steps no more: Global convergence of stochastic gradient bandits for arbitrary learning rates
von: Mei, Jincheng, et al.
Veröffentlicht: (2025)
von: Mei, Jincheng, et al.
Veröffentlicht: (2025)
Instance-dependent Stochastic Lipschitz bandit
von: Potfer, Marius, et al.
Veröffentlicht: (2026)
von: Potfer, Marius, et al.
Veröffentlicht: (2026)
Spectral bandits for smooth graph functions
von: Valko, Michal, et al.
Veröffentlicht: (2026)
von: Valko, Michal, et al.
Veröffentlicht: (2026)
Approximate information maximization for bandit games
von: Barbier-Chebbah, Alex, et al.
Veröffentlicht: (2023)
von: Barbier-Chebbah, Alex, et al.
Veröffentlicht: (2023)
Risk and optimal policies in bandit experiments
von: Adusumilli, Karun
Veröffentlicht: (2021)
von: Adusumilli, Karun
Veröffentlicht: (2021)
On the optimal regret of collaborative personalized linear bandits
von: Huang, Bruce, et al.
Veröffentlicht: (2025)
von: Huang, Bruce, et al.
Veröffentlicht: (2025)
Offline-to-online hyperparameter transfer for stochastic bandits
von: Sharma, Dravyansh, et al.
Veröffentlicht: (2025)
von: Sharma, Dravyansh, et al.
Veröffentlicht: (2025)
Revealing graph bandits for maximizing local influence
von: Carpentier, Alexandra, et al.
Veröffentlicht: (2026)
von: Carpentier, Alexandra, et al.
Veröffentlicht: (2026)
Multi-task neural networks by learned contextual inputs
von: Sandnes, Anders T., et al.
Veröffentlicht: (2023)
von: Sandnes, Anders T., et al.
Veröffentlicht: (2023)
Linear bandits with polylogarithmic minimax regret
von: Lumbreras, Josep, et al.
Veröffentlicht: (2024)
von: Lumbreras, Josep, et al.
Veröffentlicht: (2024)
Efficient kernelized bandit algorithms via exploration distributions
von: Hu, Bingshan, et al.
Veröffentlicht: (2025)
von: Hu, Bingshan, et al.
Veröffentlicht: (2025)
Leveraging priors on distribution functions for multi-arm bandits
von: Vashishtha, Sumit, et al.
Veröffentlicht: (2025)
von: Vashishtha, Sumit, et al.
Veröffentlicht: (2025)
When and why randomised exploration works (in linear bandits)
von: Abeille, Marc, et al.
Veröffentlicht: (2025)
von: Abeille, Marc, et al.
Veröffentlicht: (2025)
Lookahead identification in adversarial bandits: accuracy and memory bounds
von: Brukhim, Nataly, et al.
Veröffentlicht: (2026)
von: Brukhim, Nataly, et al.
Veröffentlicht: (2026)
Trading off rewards and errors in multi-armed bandits
von: Erraqabi, Akram, et al.
Veröffentlicht: (2026)
von: Erraqabi, Akram, et al.
Veröffentlicht: (2026)
Ensemble sampling for linear bandits: small ensembles suffice
von: Janz, David, et al.
Veröffentlicht: (2023)
von: Janz, David, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
General oracle inequalities for a penalized log-likelihood criterion based on non-stationary data
von: Aubert, Julien, et al.
Veröffentlicht: (2024) -
Spiking Neural Models for Decision-Making Tasks with Learning
von: Jaffard, Sophie, et al.
Veröffentlicht: (2025) -
CHANI: Correlation-based Hawkes Aggregation of Neurons with bio-Inspiration
von: Jaffard, Sophie, et al.
Veröffentlicht: (2024) -
Optimal cross-learning for contextual bandits with unknown context distributions
von: Schneider, Jon, et al.
Veröffentlicht: (2024) -
Quantum contextual bandits and recommender systems for quantum data
von: Brahmachari, Shrigyan, et al.
Veröffentlicht: (2023)