Enhancing Bandit Algorithms with LLMs for Time-varying User Preferences in Streaming Recommendations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shen, Chenglei, Zhan, Yi, Yu, Weijie, Zhang, Xiao, Xu, Jun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Optimal Streaming Algorithms for Multi-Armed Bandits
von: Jin, Tianyuan, et al.
Veröffentlicht: (2024)
von: Jin, Tianyuan, et al.
Veröffentlicht: (2024)
IBCB: Efficient Inverse Batched Contextual Bandit for Behavioral Evolution History
von: Xu, Yi, et al.
Veröffentlicht: (2024)
von: Xu, Yi, et al.
Veröffentlicht: (2024)
Enhancing Preference-based Linear Bandits via Human Response Time
von: Li, Shen, et al.
Veröffentlicht: (2024)
von: Li, Shen, et al.
Veröffentlicht: (2024)
COURIER: Contrastive User Intention Reconstruction for Large-Scale Visual Recommendation
von: Yang, Jia-Qi, et al.
Veröffentlicht: (2023)
von: Yang, Jia-Qi, et al.
Veröffentlicht: (2023)
A Survey of Controllable Learning: Methods and Applications in Information Retrieval
von: Shen, Chenglei, et al.
Veröffentlicht: (2024)
von: Shen, Chenglei, et al.
Veröffentlicht: (2024)
Preference-centric Bandits: Optimality of Mixtures and Regret-efficient Algorithms
von: Tatlı, Meltem, et al.
Veröffentlicht: (2025)
von: Tatlı, Meltem, et al.
Veröffentlicht: (2025)
Multi-User Contextual Cascading Bandits for Personalized Recommendation
von: Park, Jiho, et al.
Veröffentlicht: (2025)
von: Park, Jiho, et al.
Veröffentlicht: (2025)
Contextual Bandit with Herding Effects: Algorithms and Recommendation Applications
von: Xu, Luyue, et al.
Veröffentlicht: (2024)
von: Xu, Luyue, et al.
Veröffentlicht: (2024)
Tweedie Regression for Video Recommendation System
von: Zheng, Yan, et al.
Veröffentlicht: (2025)
von: Zheng, Yan, et al.
Veröffentlicht: (2025)
Latent Preference Bandits
von: Mwai, Newton, et al.
Veröffentlicht: (2025)
von: Mwai, Newton, et al.
Veröffentlicht: (2025)
When Personalization Misleads: Understanding and Mitigating Hallucinations in Personalized LLMs
von: Sun, Zhongxiang, et al.
Veröffentlicht: (2026)
von: Sun, Zhongxiang, et al.
Veröffentlicht: (2026)
GenRecEdit: Adapting Model Editing for Generative Recommendation with Cold-Start Items
von: Shen, Chenglei, et al.
Veröffentlicht: (2026)
von: Shen, Chenglei, et al.
Veröffentlicht: (2026)
Provably Efficient Multi-Objective Bandit Algorithms under Preference-Centric Customization
von: Cao, Linfeng, et al.
Veröffentlicht: (2025)
von: Cao, Linfeng, et al.
Veröffentlicht: (2025)
Dynamic User Interest Augmentation via Stream Clustering and Memory Networks in Large-Scale Recommender Systems
von: Liu, Peng, et al.
Veröffentlicht: (2024)
von: Liu, Peng, et al.
Veröffentlicht: (2024)
Adapting Job Recommendations to User Preference Drift with Behavioral-Semantic Fusion Learning
von: Han, Xiao, et al.
Veröffentlicht: (2024)
von: Han, Xiao, et al.
Veröffentlicht: (2024)
Bayesian Bandit Algorithms with Approximate Inference in Stochastic Linear Bandits
von: Huang, Ziyi, et al.
Veröffentlicht: (2024)
von: Huang, Ziyi, et al.
Veröffentlicht: (2024)
Wasserstein Distributionally Robust Policy Evaluation and Learning for Contextual Bandits
von: Shen, Yi, et al.
Veröffentlicht: (2023)
von: Shen, Yi, et al.
Veröffentlicht: (2023)
The Bandit's Blind Spot: The Critical Role of User State Representation in Recommender Systems
von: Pires, Pedro R., et al.
Veröffentlicht: (2026)
von: Pires, Pedro R., et al.
Veröffentlicht: (2026)
Queueing Matching Bandits with Preference Feedback
von: Kim, Jung-hun, et al.
Veröffentlicht: (2024)
von: Kim, Jung-hun, et al.
Veröffentlicht: (2024)
Learning to Route and Schedule LLMs from User Retrials via Contextual Queueing Bandits
von: Bae, Seoungbin, et al.
Veröffentlicht: (2026)
von: Bae, Seoungbin, et al.
Veröffentlicht: (2026)
Modeling User Preferences as Distributions for Optimal Transport-Based Cross-Domain Recommendation under Non-Overlapping Settings
von: Xiao, Ziyin, et al.
Veröffentlicht: (2025)
von: Xiao, Ziyin, et al.
Veröffentlicht: (2025)
Unlocking Reasoning Capabilities in LLMs via Reinforcement Learning Exploration
von: Deng, Wenhao, et al.
Veröffentlicht: (2025)
von: Deng, Wenhao, et al.
Veröffentlicht: (2025)
Optimal and Practical Batched Linear Bandit Algorithm
von: Yu, Sanghoon, et al.
Veröffentlicht: (2025)
von: Yu, Sanghoon, et al.
Veröffentlicht: (2025)
The Nah Bandit: Modeling User Non-compliance in Recommendation Systems
von: Zhou, Tianyue, et al.
Veröffentlicht: (2024)
von: Zhou, Tianyue, et al.
Veröffentlicht: (2024)
On the Algorithmic Bias of Aligning Large Language Models with RLHF: Preference Collapse and Matching Regularization
von: Xiao, Jiancong, et al.
Veröffentlicht: (2024)
von: Xiao, Jiancong, et al.
Veröffentlicht: (2024)
Artificial Replay: A Meta-Algorithm for Harnessing Historical Data in Bandits
von: Banerjee, Siddhartha, et al.
Veröffentlicht: (2022)
von: Banerjee, Siddhartha, et al.
Veröffentlicht: (2022)
Algorithmic Assistance with Recommendation-Dependent Preferences
von: McLaughlin, Bryce, et al.
Veröffentlicht: (2022)
von: McLaughlin, Bryce, et al.
Veröffentlicht: (2022)
BanditSpec: Adaptive Speculative Decoding via Bandit Algorithms
von: Hou, Yunlong, et al.
Veröffentlicht: (2025)
von: Hou, Yunlong, et al.
Veröffentlicht: (2025)
Linear Bandits on Ellipsoids: Minimax Optimal Algorithms
von: Zhang, Raymond, et al.
Veröffentlicht: (2025)
von: Zhang, Raymond, et al.
Veröffentlicht: (2025)
Efficient Algorithms for Logistic Contextual Slate Bandits with Bandit Feedback
von: Goyal, Tanmay, et al.
Veröffentlicht: (2025)
von: Goyal, Tanmay, et al.
Veröffentlicht: (2025)
Save, Revisit, Retain: A Scalable Framework for Enhancing User Retention in Large-Scale Recommender Systems
von: Jiang, Weijie, et al.
Veröffentlicht: (2025)
von: Jiang, Weijie, et al.
Veröffentlicht: (2025)
Calibrated Recommendations with Contextual Bandits
von: Feijer, Diego, et al.
Veröffentlicht: (2025)
von: Feijer, Diego, et al.
Veröffentlicht: (2025)
Separating and Learning Latent Confounders to Enhancing User Preferences Modeling
von: Xu, Hangtong, et al.
Veröffentlicht: (2023)
von: Xu, Hangtong, et al.
Veröffentlicht: (2023)
Efficient and Interpretable Bandit Algorithms
von: Mukherjee, Subhojyoti, et al.
Veröffentlicht: (2023)
von: Mukherjee, Subhojyoti, et al.
Veröffentlicht: (2023)
Quantum-Enhanced Neural Contextual Bandit Algorithms
von: Huang, Yuqi, et al.
Veröffentlicht: (2026)
von: Huang, Yuqi, et al.
Veröffentlicht: (2026)
Do LLMs Benefit from User and Item Embeddings in Recommendation Tasks?
von: Hossain, Mir Rayat Imtiaz, et al.
Veröffentlicht: (2026)
von: Hossain, Mir Rayat Imtiaz, et al.
Veröffentlicht: (2026)
Aligning LLMs by Predicting Preferences from User Writing Samples
von: Aroca-Ouellette, Stéphane, et al.
Veröffentlicht: (2025)
von: Aroca-Ouellette, Stéphane, et al.
Veröffentlicht: (2025)
Harm Mitigation in Recommender Systems under User Preference Dynamics
von: Chee, Jerry, et al.
Veröffentlicht: (2024)
von: Chee, Jerry, et al.
Veröffentlicht: (2024)
FLDmamba: Integrating Fourier and Laplace Transform Decomposition with Mamba for Enhanced Time Series Prediction
von: Zhang, Qianru, et al.
Veröffentlicht: (2025)
von: Zhang, Qianru, et al.
Veröffentlicht: (2025)
Efficient and Adaptive Posterior Sampling Algorithms for Bandits
von: Hu, Bingshan, et al.
Veröffentlicht: (2024)
von: Hu, Bingshan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Optimal Streaming Algorithms for Multi-Armed Bandits
von: Jin, Tianyuan, et al.
Veröffentlicht: (2024) -
IBCB: Efficient Inverse Batched Contextual Bandit for Behavioral Evolution History
von: Xu, Yi, et al.
Veröffentlicht: (2024) -
Enhancing Preference-based Linear Bandits via Human Response Time
von: Li, Shen, et al.
Veröffentlicht: (2024) -
COURIER: Contrastive User Intention Reconstruction for Large-Scale Visual Recommendation
von: Yang, Jia-Qi, et al.
Veröffentlicht: (2023) -
A Survey of Controllable Learning: Methods and Applications in Information Retrieval
von: Shen, Chenglei, et al.
Veröffentlicht: (2024)