Empirical Comparison of Forgetting Mechanisms for UCB-based Algorithms on a Data-Driven Simulation Platform
Fuente:
arXiv
Salvato in:
| Autore principale: | Chen, Minxin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Replicable Bandits with UCB based Exploration
di: Deb, Rohan, et al.
Pubblicazione: (2026)
di: Deb, Rohan, et al.
Pubblicazione: (2026)
Clus-UCB: A Near-Optimal Algorithm for Clustered Bandits
di: Gore, Aakash, et al.
Pubblicazione: (2025)
di: Gore, Aakash, et al.
Pubblicazione: (2025)
UCB-type Algorithm for Budget-Constrained Expert Learning
di: Latypov, Ilgam, et al.
Pubblicazione: (2025)
di: Latypov, Ilgam, et al.
Pubblicazione: (2025)
Provably Efficient UCB-type Algorithms For Learning Predictive State Representations
di: Huang, Ruiquan, et al.
Pubblicazione: (2023)
di: Huang, Ruiquan, et al.
Pubblicazione: (2023)
A characterization of sample adaptivity in UCB data
di: Chen, Yilun, et al.
Pubblicazione: (2025)
di: Chen, Yilun, et al.
Pubblicazione: (2025)
Cooperative Multi-Agent Graph Bandits: UCB Algorithm and Regret Analysis
di: Paschalidis, Phevos, et al.
Pubblicazione: (2024)
di: Paschalidis, Phevos, et al.
Pubblicazione: (2024)
A UCB Bandit Algorithm for General ML-Based Estimators
di: Liu, Yajing, et al.
Pubblicazione: (2026)
di: Liu, Yajing, et al.
Pubblicazione: (2026)
Federated Learning for Data Market: Shapley-UCB for Seller Selection and Incentives
di: Chen, Kongyang, et al.
Pubblicazione: (2024)
di: Chen, Kongyang, et al.
Pubblicazione: (2024)
Data-Driven Simulator for Mechanical Circulatory Support with Domain Adversarial Neural Process
di: Sun, Sophia, et al.
Pubblicazione: (2024)
di: Sun, Sophia, et al.
Pubblicazione: (2024)
On the Suboptimality of GP-UCB under Polynomial Effective Optimism
di: Wang, Wenjia, et al.
Pubblicazione: (2023)
di: Wang, Wenjia, et al.
Pubblicazione: (2023)
Be More Diverse than the Most Diverse: Optimal Mixtures of Generative Models via Mixture-UCB Bandit Algorithms
di: Rezaei, Parham, et al.
Pubblicazione: (2024)
di: Rezaei, Parham, et al.
Pubblicazione: (2024)
Truncated LinUCB for Stochastic Linear Bandits
di: Song, Yanglei, et al.
Pubblicazione: (2022)
di: Song, Yanglei, et al.
Pubblicazione: (2022)
Tractable Instances of Bilinear Maximization: Implementing LinUCB on Ellipsoids
di: Zhang, Raymond, et al.
Pubblicazione: (2025)
di: Zhang, Raymond, et al.
Pubblicazione: (2025)
UCB for Large-Scale Pure Exploration: Beyond Sub-Gaussianity
di: Li, Zaile, et al.
Pubblicazione: (2025)
di: Li, Zaile, et al.
Pubblicazione: (2025)
UCB Exploration for Fixed-Budget Bayesian Best Arm Identification
di: Zhu, Rong J. B., et al.
Pubblicazione: (2024)
di: Zhu, Rong J. B., et al.
Pubblicazione: (2024)
Efficient Implementation of LinearUCB through Algorithmic Improvements and Vector Computing Acceleration for Embedded Learning Systems
di: Angioli, Marco, et al.
Pubblicazione: (2025)
di: Angioli, Marco, et al.
Pubblicazione: (2025)
FIT to Forget: Robust Continual Unlearning for Large Language Models
di: Xu, Xiaoyu, et al.
Pubblicazione: (2026)
di: Xu, Xiaoyu, et al.
Pubblicazione: (2026)
Revisiting Social Welfare in Bandits: UCB is (Nearly) All You Need
di: Sarkar, Dhruv, et al.
Pubblicazione: (2025)
di: Sarkar, Dhruv, et al.
Pubblicazione: (2025)
UCB-driven Utility Function Search for Multi-objective Reinforcement Learning
di: Shi, Yucheng, et al.
Pubblicazione: (2024)
di: Shi, Yucheng, et al.
Pubblicazione: (2024)
DAK-UCB: Diversity-Aware Prompt Routing for LLMs and Generative Models
di: Jafari, Donya, et al.
Pubblicazione: (2026)
di: Jafari, Donya, et al.
Pubblicazione: (2026)
Variance-Aware Linear UCB with Deep Representation for Neural Contextual Bandits
di: Bui, Ha Manh, et al.
Pubblicazione: (2024)
di: Bui, Ha Manh, et al.
Pubblicazione: (2024)
Precise Asymptotics and Refined Regret of Variance-Aware UCB
di: Fan, Yingying, et al.
Pubblicazione: (2024)
di: Fan, Yingying, et al.
Pubblicazione: (2024)
An Empirical Analysis of Forgetting in Pre-trained Models with Incremental Low-Rank Updates
di: Soutif--Cormerais, Albin, et al.
Pubblicazione: (2024)
di: Soutif--Cormerais, Albin, et al.
Pubblicazione: (2024)
Comparison of Outlier Detection Algorithms on String Data
di: Maus, Philip
Pubblicazione: (2026)
di: Maus, Philip
Pubblicazione: (2026)
A Spatially Informed Gaussian Process UCB Method for Decentralized Coverage Control
di: Guidone, Gennaro, et al.
Pubblicazione: (2025)
di: Guidone, Gennaro, et al.
Pubblicazione: (2025)
Statistical Inference under Adaptive Sampling with LinUCB
di: Fan, Wei, et al.
Pubblicazione: (2025)
di: Fan, Wei, et al.
Pubblicazione: (2025)
Extended UCB Policies for Multi-armed Bandit Problems
di: Liu, Keqin, et al.
Pubblicazione: (2011)
di: Liu, Keqin, et al.
Pubblicazione: (2011)
Polynomial Regret Concentration of UCB for Non-Deterministic State Transitions
di: Cömer, Can, et al.
Pubblicazione: (2025)
di: Cömer, Can, et al.
Pubblicazione: (2025)
Reward-Based Online LLM Routing via NeuralUCB
di: Tsai, Ming-Hua, et al.
Pubblicazione: (2026)
di: Tsai, Ming-Hua, et al.
Pubblicazione: (2026)
On the Convergence of Monte Carlo UCB for Random-Length Episodic MDPs
di: Dong, Zixuan, et al.
Pubblicazione: (2022)
di: Dong, Zixuan, et al.
Pubblicazione: (2022)
ClusterUCB: Efficient Gradient-Based Data Selection for Targeted Fine-Tuning of LLMs
di: Wang, Zige, et al.
Pubblicazione: (2025)
di: Wang, Zige, et al.
Pubblicazione: (2025)
System-Aware Unlearning Algorithms: Use Lesser, Forget Faster
di: Lu, Linda, et al.
Pubblicazione: (2025)
di: Lu, Linda, et al.
Pubblicazione: (2025)
Retaining by Doing: The Role of On-Policy Data in Mitigating Forgetting
di: Chen, Howard, et al.
Pubblicazione: (2025)
di: Chen, Howard, et al.
Pubblicazione: (2025)
Minimizing UCB: a Better Local Search Strategy in Local Bayesian Optimization
di: Fan, Zheyi, et al.
Pubblicazione: (2024)
di: Fan, Zheyi, et al.
Pubblicazione: (2024)
Contrastive UCB: Provably Efficient Contrastive Self-Supervised Learning in Online Reinforcement Learning
di: Qiu, Shuang, et al.
Pubblicazione: (2022)
di: Qiu, Shuang, et al.
Pubblicazione: (2022)
Context-Free Synthetic Data Mitigates Forgetting
di: Bansal, Parikshit, et al.
Pubblicazione: (2025)
di: Bansal, Parikshit, et al.
Pubblicazione: (2025)
Privacy-Preserving UCB Decision Process Verification via zk-SNARKs
di: Jiang, Xikun, et al.
Pubblicazione: (2024)
di: Jiang, Xikun, et al.
Pubblicazione: (2024)
SCM: Sleep-Consolidated Memory with Algorithmic Forgetting for Large Language Models
di: Shinde, Saish Sachin
Pubblicazione: (2026)
di: Shinde, Saish Sachin
Pubblicazione: (2026)
AdaGrad Meets Muon: Adaptive Stepsizes for Orthogonal Updates
di: Zhang, Minxin, et al.
Pubblicazione: (2025)
di: Zhang, Minxin, et al.
Pubblicazione: (2025)
Connecting Thompson Sampling and UCB: Towards More Efficient Trade-offs Between Privacy and Regret
di: Hu, Bingshan, et al.
Pubblicazione: (2025)
di: Hu, Bingshan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Replicable Bandits with UCB based Exploration
di: Deb, Rohan, et al.
Pubblicazione: (2026) -
Clus-UCB: A Near-Optimal Algorithm for Clustered Bandits
di: Gore, Aakash, et al.
Pubblicazione: (2025) -
UCB-type Algorithm for Budget-Constrained Expert Learning
di: Latypov, Ilgam, et al.
Pubblicazione: (2025) -
Provably Efficient UCB-type Algorithms For Learning Predictive State Representations
di: Huang, Ruiquan, et al.
Pubblicazione: (2023) -
A characterization of sample adaptivity in UCB data
di: Chen, Yilun, et al.
Pubblicazione: (2025)