Two-Stage Constrained Actor-Critic for Short Video Recommendation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cai, Qingpeng, Xue, Zhenghai, Zhang, Chi, Xue, Wanqi, Liu, Shuchang, Zhan, Ruohan, Wang, Xueliang, Zuo, Tianyou, Xie, Wentao, Zheng, Dong, Jiang, Peng, Gai, Kun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AURO: Reinforcement Learning for Adaptive User Retention Optimization in Recommender Systems
von: Xue, Zhenghai, et al.
Veröffentlicht: (2023)
von: Xue, Zhenghai, et al.
Veröffentlicht: (2023)
Future Impact Decomposition in Request-level Recommendations
von: Wang, Xiaobei, et al.
Veröffentlicht: (2024)
von: Wang, Xiaobei, et al.
Veröffentlicht: (2024)
Coarse-to-fine Dynamic Uplift Modeling for Real-time Video Recommendation
von: Meng, Chang, et al.
Veröffentlicht: (2024)
von: Meng, Chang, et al.
Veröffentlicht: (2024)
Sequential Recommendation for Optimizing Both Immediate Feedback and Long-term Retention
von: Liu, Ziru, et al.
Veröffentlicht: (2024)
von: Liu, Ziru, et al.
Veröffentlicht: (2024)
Modeling User Retention through Generative Flow Networks
von: Liu, Ziru, et al.
Veröffentlicht: (2024)
von: Liu, Ziru, et al.
Veröffentlicht: (2024)
Heterogeneous Multi-treatment Uplift Modeling for Trade-off Optimization in Short-Video Recommendation
von: Zhai, Chenhao, et al.
Veröffentlicht: (2025)
von: Zhai, Chenhao, et al.
Veröffentlicht: (2025)
From Principles to Applications: A Comprehensive Survey of Discrete Tokenizers in Generation, Comprehension, Recommendation, and Information Retrieval
von: Jia, Jian, et al.
Veröffentlicht: (2025)
von: Jia, Jian, et al.
Veröffentlicht: (2025)
DLCRec: A Novel Approach for Managing Diversity in LLM-Based Recommender Systems
von: Chen, Jiaju, et al.
Veröffentlicht: (2024)
von: Chen, Jiaju, et al.
Veröffentlicht: (2024)
Reinforced Preference Optimization for Reasoning-Augmented Recommendations
von: Gao, Jingtong, et al.
Veröffentlicht: (2026)
von: Gao, Jingtong, et al.
Veröffentlicht: (2026)
Value Function Decomposition in Markov Recommendation Process
von: Wang, Xiaobei, et al.
Veröffentlicht: (2025)
von: Wang, Xiaobei, et al.
Veröffentlicht: (2025)
State Regularized Policy Optimization on Data with Dynamics Shift
von: Xue, Zhenghai, et al.
Veröffentlicht: (2023)
von: Xue, Zhenghai, et al.
Veröffentlicht: (2023)
M3oE: Multi-Domain Multi-Task Mixture-of Experts Recommendation Framework
von: Zhang, Zijian, et al.
Veröffentlicht: (2024)
von: Zhang, Zijian, et al.
Veröffentlicht: (2024)
Future-Conditioned Recommendations with Multi-Objective Controllable Decision Transformer
von: Gao, Chongming, et al.
Veröffentlicht: (2025)
von: Gao, Chongming, et al.
Veröffentlicht: (2025)
LLM-Powered User Simulator for Recommender System
von: Zhang, Zijian, et al.
Veröffentlicht: (2024)
von: Zhang, Zijian, et al.
Veröffentlicht: (2024)
Towards End-to-End Alignment of User Satisfaction via Questionnaire in Video Recommendation
von: Li, Na, et al.
Veröffentlicht: (2026)
von: Li, Na, et al.
Veröffentlicht: (2026)
An End-to-End Multi-objective Ensemble Ranking Framework for Video Recommendation
von: He, Tiantian, et al.
Veröffentlicht: (2025)
von: He, Tiantian, et al.
Veröffentlicht: (2025)
Denoising Neural Reranker for Recommender Systems
von: Mao, Wenyu, et al.
Veröffentlicht: (2025)
von: Mao, Wenyu, et al.
Veröffentlicht: (2025)
Fading to Grow: Growing Preference Ratios via Preference Fading Discrete Diffusion for Recommendation
von: Hu, Guoqing, et al.
Veröffentlicht: (2025)
von: Hu, Guoqing, et al.
Veröffentlicht: (2025)
TrackRec: Iterative Alternating Feedback with Chain-of-Thought via Preference Alignment for Recommendation
von: Xia, Yu, et al.
Veröffentlicht: (2025)
von: Xia, Yu, et al.
Veröffentlicht: (2025)
Research on the Design of a Short Video Recommendation System Based on Multimodal Information and Differential Privacy
von: Yang, Haowei, et al.
Veröffentlicht: (2025)
von: Yang, Haowei, et al.
Veröffentlicht: (2025)
R4ec: A Reasoning, Reflection, and Refinement Framework for Recommendation Systems
von: Gu, Hao, et al.
Veröffentlicht: (2025)
von: Gu, Hao, et al.
Veröffentlicht: (2025)
Relative Advantage Debiasing for Watch-Time Prediction in Short-Video Recommendation
von: Liu, Emily, et al.
Veröffentlicht: (2025)
von: Liu, Emily, et al.
Veröffentlicht: (2025)
Creator-Side Recommender System: Challenges, Designs, and Applications
von: Chen, Xiaoshuang, et al.
Veröffentlicht: (2025)
von: Chen, Xiaoshuang, et al.
Veröffentlicht: (2025)
SocRipple: A Two-Stage Framework for Cold-Start Video Recommendations
von: Jaspal, Amit, et al.
Veröffentlicht: (2025)
von: Jaspal, Amit, et al.
Veröffentlicht: (2025)
Stratified Expert Cloning for Retention-Aware Recommendation at Scale
von: Lin, Chengzhi, et al.
Veröffentlicht: (2025)
von: Lin, Chengzhi, et al.
Veröffentlicht: (2025)
UniRank: Unified List-wise Reranking via Confidence-Ordered Denoising
von: Jia, Pengyue, et al.
Veröffentlicht: (2026)
von: Jia, Pengyue, et al.
Veröffentlicht: (2026)
Who You Are Matters: Bridging Topics and Social Roles via LLM-Enhanced Logical Recommendation
von: Yu, Qing, et al.
Veröffentlicht: (2025)
von: Yu, Qing, et al.
Veröffentlicht: (2025)
PROMISE: Process Reward Models Unlock Test-Time Scaling Laws in Generative Recommendations
von: Guo, Chengcheng, et al.
Veröffentlicht: (2026)
von: Guo, Chengcheng, et al.
Veröffentlicht: (2026)
From Local Indices to Global Identifiers: Generative Reranking for Recommender Systems via Global Action Space
von: Jia, Pengyue, et al.
Veröffentlicht: (2026)
von: Jia, Pengyue, et al.
Veröffentlicht: (2026)
Full Stage Learning to Rank: A Unified Framework for Multi-Stage Systems
von: Zheng, Kai, et al.
Veröffentlicht: (2024)
von: Zheng, Kai, et al.
Veröffentlicht: (2024)
Unleashing the Native Recommendation Potential: LLM-Based Generative Recommendation via Structured Term Identifiers
von: Zhang, Zhiyang, et al.
Veröffentlicht: (2026)
von: Zhang, Zhiyang, et al.
Veröffentlicht: (2026)
MindRec: A Diffusion-driven Coarse-to-Fine Paradigm for Generative Recommendation
von: Gao, Mengyao, et al.
Veröffentlicht: (2025)
von: Gao, Mengyao, et al.
Veröffentlicht: (2025)
Hierarchical Semantic RL: Tackling the Problem of Dynamic Action Space for RL-based Recommendations
von: Wang, Minmao, et al.
Veröffentlicht: (2025)
von: Wang, Minmao, et al.
Veröffentlicht: (2025)
GEMs: Breaking the Long-Sequence Barrier in Generative Recommendation with a Multi-Stream Decoder
von: Zhou, Yu, et al.
Veröffentlicht: (2026)
von: Zhou, Yu, et al.
Veröffentlicht: (2026)
RPAF: A Reinforcement Prediction-Allocation Framework for Cache Allocation in Large-Scale Recommender Systems
von: Su, Shuo, et al.
Veröffentlicht: (2024)
von: Su, Shuo, et al.
Veröffentlicht: (2024)
Not All Videos Become Outdated: Short-Video Recommendation by Learning to Deconfound Release Interval Bias
von: Dong, Lulu, et al.
Veröffentlicht: (2024)
von: Dong, Lulu, et al.
Veröffentlicht: (2024)
Short Video Segment-level User Dynamic Interests Modeling in Personalized Recommendation
von: He, Zhiyu, et al.
Veröffentlicht: (2025)
von: He, Zhiyu, et al.
Veröffentlicht: (2025)
Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation
von: Chen, Jiaju, et al.
Veröffentlicht: (2026)
von: Chen, Jiaju, et al.
Veröffentlicht: (2026)
MISS: Multi-Modal Tree Indexing and Searching with Lifelong Sequential Behavior for Retrieval Recommendation
von: Guo, Chengcheng, et al.
Veröffentlicht: (2025)
von: Guo, Chengcheng, et al.
Veröffentlicht: (2025)
Modeling User Fatigue for Sequential Recommendation
von: Li, Nian, et al.
Veröffentlicht: (2024)
von: Li, Nian, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
AURO: Reinforcement Learning for Adaptive User Retention Optimization in Recommender Systems
von: Xue, Zhenghai, et al.
Veröffentlicht: (2023) -
Future Impact Decomposition in Request-level Recommendations
von: Wang, Xiaobei, et al.
Veröffentlicht: (2024) -
Coarse-to-fine Dynamic Uplift Modeling for Real-time Video Recommendation
von: Meng, Chang, et al.
Veröffentlicht: (2024) -
Sequential Recommendation for Optimizing Both Immediate Feedback and Long-term Retention
von: Liu, Ziru, et al.
Veröffentlicht: (2024) -
Modeling User Retention through Generative Flow Networks
von: Liu, Ziru, et al.
Veröffentlicht: (2024)