Saved in:
| Main Authors: | Guo, Zhanqiu, Wang, Wayne |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2410.09781 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Meta Clustering of Neural Bandits
by: Ban, Yikun, et al.
Published: (2024)
by: Ban, Yikun, et al.
Published: (2024)
Facet-Aware Multi-Head Mixture-of-Experts Model for Sequential Recommendation
by: Liu, Mingrui, et al.
Published: (2024)
by: Liu, Mingrui, et al.
Published: (2024)
Deep Uncertainty-Based Explore for Index Construction and Retrieval in Recommendation System
by: Jiang, Xin, et al.
Published: (2024)
by: Jiang, Xin, et al.
Published: (2024)
RouterKT: Mixture-of-Experts for Knowledge Tracing
by: Liao, Han, et al.
Published: (2025)
by: Liao, Han, et al.
Published: (2025)
Whittle Index Learning Algorithms for Restless Bandits with Constant Stepsizes
by: Mittal, Vishesh, et al.
Published: (2024)
by: Mittal, Vishesh, et al.
Published: (2024)
Modeling Attrition in Recommender Systems with Departing Bandits
by: Ben-Porat, Omer, et al.
Published: (2022)
by: Ben-Porat, Omer, et al.
Published: (2022)
Mixture-of-PageRanks: Replacing Long-Context with Real-Time, Sparse GraphRAG
by: Alonso, Nicholas, et al.
Published: (2024)
by: Alonso, Nicholas, et al.
Published: (2024)
Calibrated Recommendations with Contextual Bandits
by: Feijer, Diego, et al.
Published: (2025)
by: Feijer, Diego, et al.
Published: (2025)
Influence Maximization via Graph Neural Bandits
by: Feng, Yuting, et al.
Published: (2024)
by: Feng, Yuting, et al.
Published: (2024)
Optimal Baseline Corrections for Off-Policy Contextual Bandits
by: Gupta, Shashank, et al.
Published: (2024)
by: Gupta, Shashank, et al.
Published: (2024)
Deep Adaptive Interest Network: Personalized Recommendation with Context-Aware Learning
by: Huang, Shuaishuai, et al.
Published: (2024)
by: Huang, Shuaishuai, et al.
Published: (2024)
EasyRL4Rec: An Easy-to-use Library for Reinforcement Learning Based Recommender Systems
by: Yu, Yuanqing, et al.
Published: (2024)
by: Yu, Yuanqing, et al.
Published: (2024)
General Agentic Memory Via Deep Research
by: Yan, B. Y., et al.
Published: (2025)
by: Yan, B. Y., et al.
Published: (2025)
Adaptive Conditional Expert Selection Network for Multi-domain Recommendation
by: Dong, Kuiyao, et al.
Published: (2024)
by: Dong, Kuiyao, et al.
Published: (2024)
Contextual Bandits in Payment Processing: Non-uniform Exploration and Supervised Learning
by: Vangara, Akhila, et al.
Published: (2024)
by: Vangara, Akhila, et al.
Published: (2024)
Causal Feature Selection Method for Contextual Multi-Armed Bandits in Recommender System
by: Zhao, Zhenyu, et al.
Published: (2024)
by: Zhao, Zhenyu, et al.
Published: (2024)
The Bandit's Blind Spot: The Critical Role of User State Representation in Recommender Systems
by: Pires, Pedro R., et al.
Published: (2026)
by: Pires, Pedro R., et al.
Published: (2026)
Exploitation Over Exploration: Unmasking the Bias in Linear Bandit Recommender Offline Evaluation
by: Pires, Pedro R., et al.
Published: (2025)
by: Pires, Pedro R., et al.
Published: (2025)
Enhancing Talent Search Ranking with Role-Aware Expert Mixtures and LLM-based Fine-Grained Job Descriptions
by: Li, Jihang, et al.
Published: (2025)
by: Li, Jihang, et al.
Published: (2025)
BiCoRec: Bias-Mitigated Context-Aware Sequential Recommendation Model
by: Muthivhi, Mufhumudzi, et al.
Published: (2025)
by: Muthivhi, Mufhumudzi, et al.
Published: (2025)
EvidenceRL: Reinforcing Evidence Consistency for Trustworthy Language Models
by: Tamo, J. Ben, et al.
Published: (2026)
by: Tamo, J. Ben, et al.
Published: (2026)
Vectorized Context-Aware Embeddings for GAT-Based Collaborative Filtering
by: Ebrat, Danial, et al.
Published: (2025)
by: Ebrat, Danial, et al.
Published: (2025)
EnhancedRL: An Enhanced-State Reinforcement Learning Algorithm for Multi-Task Fusion in Recommender Systems
by: Liu, Peng, et al.
Published: (2024)
by: Liu, Peng, et al.
Published: (2024)
SUBER: An RL Environment with Simulated Human Behavior for Recommender Systems
by: Corecco, Nathan, et al.
Published: (2024)
by: Corecco, Nathan, et al.
Published: (2024)
EulerFormer: Sequential User Behavior Modeling with Complex Vector Attention
by: Tian, Zhen, et al.
Published: (2024)
by: Tian, Zhen, et al.
Published: (2024)
Ranking In Generalized Linear Bandits
by: Shidani, Amitis, et al.
Published: (2022)
by: Shidani, Amitis, et al.
Published: (2022)
Contextual Bandit with Herding Effects: Algorithms and Recommendation Applications
by: Xu, Luyue, et al.
Published: (2024)
by: Xu, Luyue, et al.
Published: (2024)
A Production-Ready RL Framework for Personalized Utility Tuning with Pareto Sweeping in Pinterest Recommender Systems
by: Zhou, Yichu, et al.
Published: (2026)
by: Zhou, Yichu, et al.
Published: (2026)
VIBE: Vector Index Benchmark for Embeddings
by: Jääsaari, Elias, et al.
Published: (2025)
by: Jääsaari, Elias, et al.
Published: (2025)
UnifiedRL: A Reinforcement Learning Algorithm Tailored for Multi-Task Fusion in Large-Scale Recommender Systems
by: Liu, Peng, et al.
Published: (2024)
by: Liu, Peng, et al.
Published: (2024)
HoME: Hierarchy of Multi-Gate Experts for Multi-Task Learning at Kuaishou
by: Wang, Xu, et al.
Published: (2024)
by: Wang, Xu, et al.
Published: (2024)
EBaReT: Expert-guided Bag Reward Transformer for Auto Bidding
by: Li, Kaiyuan, et al.
Published: (2025)
by: Li, Kaiyuan, et al.
Published: (2025)
Expert-Guided Diffusion Planner for Auto-Bidding
by: Peng, Yunshan, et al.
Published: (2025)
by: Peng, Yunshan, et al.
Published: (2025)
Shielded RecRL: Explanation Generation for Recommender Systems without Ranking Degradation
by: Tiwari, Ansh, et al.
Published: (2025)
by: Tiwari, Ansh, et al.
Published: (2025)
Efficient, Property-Aligned Fan-Out Retrieval via RL-Compiled Diffusion
by: Jiang, Pengcheng, et al.
Published: (2026)
by: Jiang, Pengcheng, et al.
Published: (2026)
RL-MPCA: A Reinforcement Learning Based Multi-Phase Computation Allocation Approach for Recommender Systems
by: Zhou, Jiahong, et al.
Published: (2023)
by: Zhou, Jiahong, et al.
Published: (2023)
Prompt Optimization with Logged Bandit Data
by: Kiyohara, Haruka, et al.
Published: (2025)
by: Kiyohara, Haruka, et al.
Published: (2025)
ContextGNN goes to Elliot: Towards Benchmarking Relational Deep Learning for Static Link Prediction (aka Personalized Item Recommendation)
by: Ariza-Casabona, Alejandro, et al.
Published: (2025)
by: Ariza-Casabona, Alejandro, et al.
Published: (2025)
An Integrated Data Processing Framework for Pretraining Foundation Models
by: Sun, Yiding, et al.
Published: (2024)
by: Sun, Yiding, et al.
Published: (2024)
Restless Linear Bandits
by: Khaleghi, Azadeh
Published: (2024)
by: Khaleghi, Azadeh
Published: (2024)
Similar Items
-
Meta Clustering of Neural Bandits
by: Ban, Yikun, et al.
Published: (2024) -
Facet-Aware Multi-Head Mixture-of-Experts Model for Sequential Recommendation
by: Liu, Mingrui, et al.
Published: (2024) -
Deep Uncertainty-Based Explore for Index Construction and Retrieval in Recommendation System
by: Jiang, Xin, et al.
Published: (2024) -
RouterKT: Mixture-of-Experts for Knowledge Tracing
by: Liao, Han, et al.
Published: (2025) -
Whittle Index Learning Algorithms for Restless Bandits with Constant Stepsizes
by: Mittal, Vishesh, et al.
Published: (2024)