Reinforced Preference Optimization for Reasoning-Augmented Recommendations
Fuente:
arXiv
Saved in:
| Main Authors: | Gao, Jingtong, Song, Zeyu, Lu, Chi, Li, Xiaopeng, Xu, Derong, Wang, Maolin, Jiang, Peng, Gai, Kun, Cai, Qingpeng, Zhao, Xiangyu |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
From Principles to Applications: A Comprehensive Survey of Discrete Tokenizers in Generation, Comprehension, Recommendation, and Information Retrieval
by: Jia, Jian, et al.
Published: (2025)
by: Jia, Jian, et al.
Published: (2025)
AURO: Reinforcement Learning for Adaptive User Retention Optimization in Recommender Systems
by: Xue, Zhenghai, et al.
Published: (2023)
by: Xue, Zhenghai, et al.
Published: (2023)
Sequential Recommendation for Optimizing Both Immediate Feedback and Long-term Retention
by: Liu, Ziru, et al.
Published: (2024)
by: Liu, Ziru, et al.
Published: (2024)
TrackRec: Iterative Alternating Feedback with Chain-of-Thought via Preference Alignment for Recommendation
by: Xia, Yu, et al.
Published: (2025)
by: Xia, Yu, et al.
Published: (2025)
Generative Auto-Bidding with Value-Guided Explorations
by: Gao, Jingtong, et al.
Published: (2025)
by: Gao, Jingtong, et al.
Published: (2025)
SampleLLM: Optimizing Tabular Data Synthesis in Recommendations
by: Gao, Jingtong, et al.
Published: (2025)
by: Gao, Jingtong, et al.
Published: (2025)
R4ec: A Reasoning, Reflection, and Refinement Framework for Recommendation Systems
by: Gu, Hao, et al.
Published: (2025)
by: Gu, Hao, et al.
Published: (2025)
Reinforced Preference Optimization for Recommendation
by: Tan, Junfei, et al.
Published: (2025)
by: Tan, Junfei, et al.
Published: (2025)
M3oE: Multi-Domain Multi-Task Mixture-of Experts Recommendation Framework
by: Zhang, Zijian, et al.
Published: (2024)
by: Zhang, Zijian, et al.
Published: (2024)
SSDRec: Self-Augmented Sequence Denoising for Sequential Recommendation
by: Zhang, Chi, et al.
Published: (2024)
by: Zhang, Chi, et al.
Published: (2024)
GLINT-RU: Gated Lightweight Intelligent Recurrent Units for Sequential Recommender Systems
by: Zhang, Sheng, et al.
Published: (2024)
by: Zhang, Sheng, et al.
Published: (2024)
Future Impact Decomposition in Request-level Recommendations
by: Wang, Xiaobei, et al.
Published: (2024)
by: Wang, Xiaobei, et al.
Published: (2024)
Scenario-Wise Rec: A Multi-Scenario Recommendation Benchmark
by: Li, Xiaopeng, et al.
Published: (2024)
by: Li, Xiaopeng, et al.
Published: (2024)
Joint Modeling in Recommendations: A Survey
by: Zhao, Xiangyu, et al.
Published: (2025)
by: Zhao, Xiangyu, et al.
Published: (2025)
DLCRec: A Novel Approach for Managing Diversity in LLM-Based Recommender Systems
by: Chen, Jiaju, et al.
Published: (2024)
by: Chen, Jiaju, et al.
Published: (2024)
Two-Stage Constrained Actor-Critic for Short Video Recommendation
by: Cai, Qingpeng, et al.
Published: (2023)
by: Cai, Qingpeng, et al.
Published: (2023)
Empowering Denoising Sequential Recommendation with Large Language Model Embeddings
by: Wu, Tongzhou, et al.
Published: (2025)
by: Wu, Tongzhou, et al.
Published: (2025)
LLM-Powered User Simulator for Recommender System
by: Zhang, Zijian, et al.
Published: (2024)
by: Zhang, Zijian, et al.
Published: (2024)
MindRec: A Diffusion-driven Coarse-to-Fine Paradigm for Generative Recommendation
by: Gao, Mengyao, et al.
Published: (2025)
by: Gao, Mengyao, et al.
Published: (2025)
PRISM: Purified Representation and Integrated Semantic Modeling for Generative Sequential Recommendation
by: Fang, Dengzhao, et al.
Published: (2026)
by: Fang, Dengzhao, et al.
Published: (2026)
Measure Domain's Gap: A Similar Domain Selection Principle for Multi-Domain Recommendation
by: Wen, Yi, et al.
Published: (2025)
by: Wen, Yi, et al.
Published: (2025)
HiD-VAE: Interpretable Generative Recommendation via Hierarchical and Disentangled Semantic IDs
by: Fang, Dengzhao, et al.
Published: (2025)
by: Fang, Dengzhao, et al.
Published: (2025)
LinRec: Linear Attention Mechanism for Long-term Sequential Recommender Systems
by: Liu, Langming, et al.
Published: (2024)
by: Liu, Langming, et al.
Published: (2024)
Detecting Miscitation on the Scholarly Web through LLM-Augmented Text-Rich Graph Learning
by: Wu, Huidong, et al.
Published: (2026)
by: Wu, Huidong, et al.
Published: (2026)
Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation
by: Chen, Jiaju, et al.
Published: (2026)
by: Chen, Jiaju, et al.
Published: (2026)
Hierarchical Semantic RL: Tackling the Problem of Dynamic Action Space for RL-based Recommendations
by: Wang, Minmao, et al.
Published: (2025)
by: Wang, Minmao, et al.
Published: (2025)
Future-Conditioned Recommendations with Multi-Objective Controllable Decision Transformer
by: Gao, Chongming, et al.
Published: (2025)
by: Gao, Chongming, et al.
Published: (2025)
To Search or Not to Search: Aligning the Decision Boundary of Deep Search Agents via Causal Intervention
by: Zhang, Wenlin, et al.
Published: (2026)
by: Zhang, Wenlin, et al.
Published: (2026)
Navigate the Unknown: Enhancing LLM Reasoning with Intrinsic Motivation Guided Exploration
by: Gao, Jingtong, et al.
Published: (2025)
by: Gao, Jingtong, et al.
Published: (2025)
Multimodal Recommender Systems: A Survey
by: Liu, Qidong, et al.
Published: (2023)
by: Liu, Qidong, et al.
Published: (2023)
Fading to Grow: Growing Preference Ratios via Preference Fading Discrete Diffusion for Recommendation
by: Hu, Guoqing, et al.
Published: (2025)
by: Hu, Guoqing, et al.
Published: (2025)
Modeling User Fatigue for Sequential Recommendation
by: Li, Nian, et al.
Published: (2024)
by: Li, Nian, et al.
Published: (2024)
Structured Spectral Reasoning for Frequency-Adaptive Multimodal Recommendation
by: Yang, Wei, et al.
Published: (2025)
by: Yang, Wei, et al.
Published: (2025)
Enhancing Interpretability and Effectiveness in Recommendation with Numerical Features via Learning to Contrast the Counterfactual samples
by: Xu, Xiaoxiao, et al.
Published: (2025)
by: Xu, Xiaoxiao, et al.
Published: (2025)
LLM4Rerank: LLM-based Auto-Reranking Framework for Recommendations
by: Gao, Jingtong, et al.
Published: (2024)
by: Gao, Jingtong, et al.
Published: (2024)
Process vs. Outcome Reward: Which is Better for Agentic RAG Reinforcement Learning
by: Zhang, Wenlin, et al.
Published: (2025)
by: Zhang, Wenlin, et al.
Published: (2025)
MGFRec: Towards Reinforced Reasoning Recommendation with Multiple Groundings and Feedback
by: Cai, Shihao, et al.
Published: (2025)
by: Cai, Shihao, et al.
Published: (2025)
RALLRec+: Retrieval Augmented Large Language Model Recommendation with Reasoning
by: Luo, Sichun, et al.
Published: (2025)
by: Luo, Sichun, et al.
Published: (2025)
BiVRec: Bidirectional View-based Multimodal Sequential Recommendation
by: Hu, Jiaxi, et al.
Published: (2024)
by: Hu, Jiaxi, et al.
Published: (2024)
SPARK: Adaptive Low-Rank Knowledge Graph Modeling in Hybrid Geometric Spaces for Recommendation
by: Wang, Binhao, et al.
Published: (2025)
by: Wang, Binhao, et al.
Published: (2025)
Similar Items
-
From Principles to Applications: A Comprehensive Survey of Discrete Tokenizers in Generation, Comprehension, Recommendation, and Information Retrieval
by: Jia, Jian, et al.
Published: (2025) -
AURO: Reinforcement Learning for Adaptive User Retention Optimization in Recommender Systems
by: Xue, Zhenghai, et al.
Published: (2023) -
Sequential Recommendation for Optimizing Both Immediate Feedback and Long-term Retention
by: Liu, Ziru, et al.
Published: (2024) -
TrackRec: Iterative Alternating Feedback with Chain-of-Thought via Preference Alignment for Recommendation
by: Xia, Yu, et al.
Published: (2025) -
Generative Auto-Bidding with Value-Guided Explorations
by: Gao, Jingtong, et al.
Published: (2025)