Speed Up the Cold-Start Learning in Two-Sided Bandits with Many Arms
Fuente:
arXiv
Guardado en:
| Autores principales: | Bayati, Mohsen, Cao, Junyu, Chen, Wanning |
|---|---|
| Formato: | Preprint |
| Publicado: |
2022
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
The Unreasonable Effectiveness of Greedy Algorithms in Multi-Armed Bandit with Many Arms
por: Bayati, Mohsen, et al.
Publicado: (2020)
por: Bayati, Mohsen, et al.
Publicado: (2020)
Annealed Softmax Greedy in Many-Armed Bayesian Bandits
por: Overman, William, et al.
Publicado: (2026)
por: Overman, William, et al.
Publicado: (2026)
A Probabilistic Approach for Model Alignment with Human Comparisons
por: Cao, Junyu, et al.
Publicado: (2024)
por: Cao, Junyu, et al.
Publicado: (2024)
Geometry-Aware Approaches for Balancing Performance and Theoretical Guarantees in Linear Bandits
por: Luo, Yuwei, et al.
Publicado: (2023)
por: Luo, Yuwei, et al.
Publicado: (2023)
Quick-Draw Bandits: Quickly Optimizing in Nonstationary Environments with Extremely Many Arms
por: Everett, Derek, et al.
Publicado: (2025)
por: Everett, Derek, et al.
Publicado: (2025)
The Oversight Game: Learning to Cooperatively Balance an AI Agent's Safety and Autonomy
por: Overman, William, et al.
Publicado: (2025)
por: Overman, William, et al.
Publicado: (2025)
Graph Feedback Bandits with Similar Arms
por: Qi, Han, et al.
Publicado: (2024)
por: Qi, Han, et al.
Publicado: (2024)
Causal Message Passing for Experiments with Unknown and General Network Interference
por: Shirani, Sadegh, et al.
Publicado: (2023)
por: Shirani, Sadegh, et al.
Publicado: (2023)
HR-Bandit: Human-AI Collaborated Linear Recourse Bandit
por: Cao, Junyu, et al.
Publicado: (2024)
por: Cao, Junyu, et al.
Publicado: (2024)
Contextual Combinatorial Bandits with Probabilistically Triggered Arms
por: Liu, Xutong, et al.
Publicado: (2023)
por: Liu, Xutong, et al.
Publicado: (2023)
Hybrid Combinatorial Multi-armed Bandits with Probabilistically Triggered Arms
por: Zhou, Kongchang, et al.
Publicado: (2025)
por: Zhou, Kongchang, et al.
Publicado: (2025)
On Evolution-Based Models for Experimentation Under Interference
por: Shirani, Sadegh, et al.
Publicado: (2025)
por: Shirani, Sadegh, et al.
Publicado: (2025)
User-Adaptive Meta-Learning for Cold-Start Medication Recommendation with Uncertainty Filtering
por: Moghaddam, Arya Hadizadeh, et al.
Publicado: (2026)
por: Moghaddam, Arya Hadizadeh, et al.
Publicado: (2026)
On-line Learning in Tree MDPs by Treating Policies as Bandit Arms
por: Shah, Anvay, et al.
Publicado: (2026)
por: Shah, Anvay, et al.
Publicado: (2026)
Shallow AutoEncoding Recommender with Cold Start Handling via Side Features
por: Cui, Edward DongBo, et al.
Publicado: (2025)
por: Cui, Edward DongBo, et al.
Publicado: (2025)
Causal Effects with Unobserved Unit Types in Interacting Human-AI Systems
por: Overman, William, et al.
Publicado: (2026)
por: Overman, William, et al.
Publicado: (2026)
Batch-Size Independent Regret Bounds for Combinatorial Semi-Bandits with Probabilistically Triggered Arms or Independent Arms
por: Liu, Xutong, et al.
Publicado: (2022)
por: Liu, Xutong, et al.
Publicado: (2022)
Optimizing Online Advertising with Multi-Armed Bandits: Mitigating the Cold Start Problem under Auction Dynamics
por: Soboleva, Anastasiia, et al.
Publicado: (2025)
por: Soboleva, Anastasiia, et al.
Publicado: (2025)
Blessings of Multiple Good Arms in Multi-Objective Linear Bandits
por: Ann, Heesang, et al.
Publicado: (2026)
por: Ann, Heesang, et al.
Publicado: (2026)
Graph Feedback Bandits on Similar Arms: With and Without Graph Structures
por: Qi, Han, et al.
Publicado: (2025)
por: Qi, Han, et al.
Publicado: (2025)
To Start Up a Start-Up$-$Embedding Strategic Demand Development in Operational On-Demand Fulfillment via Reinforcement Learning with Information Shaping
por: Chen, Xinwei, et al.
Publicado: (2025)
por: Chen, Xinwei, et al.
Publicado: (2025)
Aligning Model Properties via Conformal Risk Control
por: Overman, William, et al.
Publicado: (2024)
por: Overman, William, et al.
Publicado: (2024)
Stochastic Graph Bandit Learning with Side-Observations
por: Gong, Xueping, et al.
Publicado: (2023)
por: Gong, Xueping, et al.
Publicado: (2023)
Jump Start or False Start? A Theoretical and Empirical Evaluation of LLM-initialized Bandits
por: Bayley, Adam, et al.
Publicado: (2026)
por: Bayley, Adam, et al.
Publicado: (2026)
Post Launch Evaluation of Policies in a High-Dimensional Setting
por: Nassiri, Shima, et al.
Publicado: (2024)
por: Nassiri, Shima, et al.
Publicado: (2024)
Deconfounded Warm-Start Thompson Sampling with Applications to Precision Medicine
por: Jaiswal, Prateek, et al.
Publicado: (2025)
por: Jaiswal, Prateek, et al.
Publicado: (2025)
Match Made with Matrix Completion: Efficient Learning under Matching Interference
por: Tang, Zhiyuan, et al.
Publicado: (2026)
por: Tang, Zhiyuan, et al.
Publicado: (2026)
Cold-Start Active Preference Learning in Socio-Economic Domains
por: Fayaz-Bakhsh, Mojtaba, et al.
Publicado: (2025)
por: Fayaz-Bakhsh, Mojtaba, et al.
Publicado: (2025)
Meet Me at the Arm: The Cooperative Multi-Armed Bandits Problem with Shareable Arms
por: Hu, Xinyi, et al.
Publicado: (2025)
por: Hu, Xinyi, et al.
Publicado: (2025)
Jump Starting Bandits with LLM-Generated Prior Knowledge
por: Alamdari, Parand A., et al.
Publicado: (2024)
por: Alamdari, Parand A., et al.
Publicado: (2024)
SPARC: Spectral Architectures Tackling the Cold-Start Problem in Graph Learning
por: Jacobs, Yahel, et al.
Publicado: (2024)
por: Jacobs, Yahel, et al.
Publicado: (2024)
Estimating Total Effects in Bipartite Experiments with Spillovers and Partial Eligibility
por: Tan, Albert, et al.
Publicado: (2025)
por: Tan, Albert, et al.
Publicado: (2025)
Dynamic Matching Bandit For Two-Sided Online Markets
por: Li, Yuantong, et al.
Publicado: (2022)
por: Li, Yuantong, et al.
Publicado: (2022)
Contrastive Learning for Cold Start Recommendation with Adaptive Feature Fusion
por: Hu, Jiacheng, et al.
Publicado: (2025)
por: Hu, Jiacheng, et al.
Publicado: (2025)
Meta-Learning for Cold-Start Personalization in Prompt-Tuned LLMs
por: Zhao, Yushang, et al.
Publicado: (2025)
por: Zhao, Yushang, et al.
Publicado: (2025)
LLMs for Cold-Start Cutting Plane Separator Configuration
por: Lawless, Connor, et al.
Publicado: (2024)
por: Lawless, Connor, et al.
Publicado: (2024)
Learning to Route LLMs from Bandit Feedback: One Policy, Many Trade-offs
por: Wei, Wang, et al.
Publicado: (2025)
por: Wei, Wang, et al.
Publicado: (2025)
Asymptotically-Optimal Gaussian Bandits with Side Observations
por: Atsidakou, Alexia, et al.
Publicado: (2025)
por: Atsidakou, Alexia, et al.
Publicado: (2025)
Addressing Cold Start For next-article Recommendation
por: Elgohary, Omar, et al.
Publicado: (2025)
por: Elgohary, Omar, et al.
Publicado: (2025)
A Novel Two-Step Fine-Tuning Pipeline for Cold-Start Active Learning in Text Classification Tasks
por: Belém, Fabiano, et al.
Publicado: (2024)
por: Belém, Fabiano, et al.
Publicado: (2024)
Ejemplares similares
-
The Unreasonable Effectiveness of Greedy Algorithms in Multi-Armed Bandit with Many Arms
por: Bayati, Mohsen, et al.
Publicado: (2020) -
Annealed Softmax Greedy in Many-Armed Bayesian Bandits
por: Overman, William, et al.
Publicado: (2026) -
A Probabilistic Approach for Model Alignment with Human Comparisons
por: Cao, Junyu, et al.
Publicado: (2024) -
Geometry-Aware Approaches for Balancing Performance and Theoretical Guarantees in Linear Bandits
por: Luo, Yuwei, et al.
Publicado: (2023) -
Quick-Draw Bandits: Quickly Optimizing in Nonstationary Environments with Extremely Many Arms
por: Everett, Derek, et al.
Publicado: (2025)