Spend Search Where It Pays: Value-Guided Structured Sampling and Optimization for Generative Recommendation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jiang, Jie, Huang, Yangru, Wang, Zeyu, Wang, Changping, Xiong, Yuling, Zhang, Jun, Yu, Huan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Breaking the Curse of Repulsion: Optimistic Distributionally Robust Policy Optimization for Off-Policy Generative Recommendation
von: Jiang, Jie, et al.
Veröffentlicht: (2026)
von: Jiang, Jie, et al.
Veröffentlicht: (2026)
S-GRec: Personalized Semantic-Aware Generative Recommendation with Asymmetric Advantage
von: Jiang, Jie, et al.
Veröffentlicht: (2026)
von: Jiang, Jie, et al.
Veröffentlicht: (2026)
Spend Less, Reason Better: Budget-Aware Value Tree Search for LLM Agents
von: Li, Yushu, et al.
Veröffentlicht: (2026)
von: Li, Yushu, et al.
Veröffentlicht: (2026)
Adaptive Discovering and Merging for Incremental Novel Class Discovery
von: Chen, Guangyao, et al.
Veröffentlicht: (2024)
von: Chen, Guangyao, et al.
Veröffentlicht: (2024)
Large Language Model for Verilog Generation with Code-Structure-Guided Reinforcement Learning
von: Wang, Ning, et al.
Veröffentlicht: (2024)
von: Wang, Ning, et al.
Veröffentlicht: (2024)
Where and What: Reasoning Dynamic and Implicit Preferences in Situated Conversational Recommendation
von: Lin, Dongding, et al.
Veröffentlicht: (2026)
von: Lin, Dongding, et al.
Veröffentlicht: (2026)
Bilevel Optimization of Agent Skills via Monte Carlo Tree Search
von: Huang, Chenyi, et al.
Veröffentlicht: (2026)
von: Huang, Chenyi, et al.
Veröffentlicht: (2026)
IA-GCN: Interactive Graph Convolutional Network for Recommendation
von: Zhang, Yinan, et al.
Veröffentlicht: (2022)
von: Zhang, Yinan, et al.
Veröffentlicht: (2022)
Pay Attention to What and Where? Interpretable Feature Extractor in Vision-based Deep Reinforcement Learning
von: Pham, Tien, et al.
Veröffentlicht: (2025)
von: Pham, Tien, et al.
Veröffentlicht: (2025)
Structuring Value Representations via Geometric Coherence in Markov Decision Processes
von: Zhang, Zuyuan, et al.
Veröffentlicht: (2026)
von: Zhang, Zuyuan, et al.
Veröffentlicht: (2026)
AnchorDrive: LLM Scenario Rollout with Anchor-Guided Diffusion Regeneration for Safety-Critical Scenario Generation
von: Jiang, Zhulin, et al.
Veröffentlicht: (2026)
von: Jiang, Zhulin, et al.
Veröffentlicht: (2026)
OptiTree: Hierarchical Thoughts Generation with Tree Search for LLM Optimization Modeling
von: Liu, Haoyang, et al.
Veröffentlicht: (2025)
von: Liu, Haoyang, et al.
Veröffentlicht: (2025)
Coarse-Guided Visual Generation via Weighted h-Transform Sampling
von: Wang, Yanghao, et al.
Veröffentlicht: (2026)
von: Wang, Yanghao, et al.
Veröffentlicht: (2026)
Renormalization Group Guided Tensor Network Structure Search
von: Wang, Maolin, et al.
Veröffentlicht: (2025)
von: Wang, Maolin, et al.
Veröffentlicht: (2025)
Value-Guided Search for Efficient Chain-of-Thought Reasoning
von: Wang, Kaiwen, et al.
Veröffentlicht: (2025)
von: Wang, Kaiwen, et al.
Veröffentlicht: (2025)
FASTER: Value-Guided Sampling for Fast RL
von: Dong, Perry, et al.
Veröffentlicht: (2026)
von: Dong, Perry, et al.
Veröffentlicht: (2026)
Collaborative Expert LLMs Guided Multi-Objective Molecular Optimization
von: Yu, Jiajun, et al.
Veröffentlicht: (2025)
von: Yu, Jiajun, et al.
Veröffentlicht: (2025)
Rethinking the Value of Agent-Generated Tests for LLM-Based Software Engineering Agents
von: Chen, Zhi, et al.
Veröffentlicht: (2026)
von: Chen, Zhi, et al.
Veröffentlicht: (2026)
Where to Search: Measure the Prior-Structured Search Space of LLM Agents
von: Song, Zhuo-Yang
Veröffentlicht: (2025)
von: Song, Zhuo-Yang
Veröffentlicht: (2025)
GraphEdit: Large Language Models for Graph Structure Learning
von: Guo, Zirui, et al.
Veröffentlicht: (2024)
von: Guo, Zirui, et al.
Veröffentlicht: (2024)
Chunks as Arms: Multi-Armed Bandit-Guided Sampling for Long-Context LLM Preference Optimization
von: Duan, Shaohua, et al.
Veröffentlicht: (2025)
von: Duan, Shaohua, et al.
Veröffentlicht: (2025)
Role-Augmented Intent-Driven Generative Search Engine Optimization
von: Chen, Xiaolu, et al.
Veröffentlicht: (2025)
von: Chen, Xiaolu, et al.
Veröffentlicht: (2025)
Structural Entropy Guided Probabilistic Coding
von: Huang, Xiang, et al.
Veröffentlicht: (2024)
von: Huang, Xiang, et al.
Veröffentlicht: (2024)
Sample-Efficient Reinforcement Learning with Temporal Logic Objectives: Leveraging the Task Specification to Guide Exploration
von: Kantaros, Yiannis, et al.
Veröffentlicht: (2024)
von: Kantaros, Yiannis, et al.
Veröffentlicht: (2024)
Bayesian-Guided Diversity in Sequential Sampling for Recommender Systems
von: Bederina, Hiba, et al.
Veröffentlicht: (2025)
von: Bederina, Hiba, et al.
Veröffentlicht: (2025)
Improving LLM Interpretability and Performance via Guided Embedding Refinement for Sequential Recommendation
von: Jia, Nanshan, et al.
Veröffentlicht: (2025)
von: Jia, Nanshan, et al.
Veröffentlicht: (2025)
AttentionRAG: Attention-Guided Context Pruning in Retrieval-Augmented Generation
von: Fang, Yixiong, et al.
Veröffentlicht: (2025)
von: Fang, Yixiong, et al.
Veröffentlicht: (2025)
Domain-Aware Cross-Attention for Cross-domain Recommendation
von: Luo, Yuhao, et al.
Veröffentlicht: (2024)
von: Luo, Yuhao, et al.
Veröffentlicht: (2024)
Enhancing Conversational Recommender Systems with Tree-Structured Knowledge and Pretrained Language Models
von: Ren, Yongwen, et al.
Veröffentlicht: (2025)
von: Ren, Yongwen, et al.
Veröffentlicht: (2025)
Where to Move Next: Zero-shot Generalization of LLMs for Next POI Recommendation
von: Feng, Shanshan, et al.
Veröffentlicht: (2024)
von: Feng, Shanshan, et al.
Veröffentlicht: (2024)
Hierarchical Optimization via LLM-Guided Objective Evolution for Mobility-on-Demand Systems
von: Zhang, Yi, et al.
Veröffentlicht: (2025)
von: Zhang, Yi, et al.
Veröffentlicht: (2025)
Multi-objective Aligned Bidword Generation Model for E-commerce Search Advertising
von: Liu, Zhenhui, et al.
Veröffentlicht: (2025)
von: Liu, Zhenhui, et al.
Veröffentlicht: (2025)
GDBA Revisited: Unleashing the Power of Guided Local Search for Distributed Constraint Optimization
von: Deng, Yanchen, et al.
Veröffentlicht: (2025)
von: Deng, Yanchen, et al.
Veröffentlicht: (2025)
Robust Search with Uncertainty-Aware Value Models for Language Model Reasoning
von: Yu, Fei, et al.
Veröffentlicht: (2025)
von: Yu, Fei, et al.
Veröffentlicht: (2025)
VC-Soup: Value-Consistency Guided Multi-Value Alignment for Large Language Models
von: Xu, Hefei, et al.
Veröffentlicht: (2026)
von: Xu, Hefei, et al.
Veröffentlicht: (2026)
Coordinating Search-Informed Reasoning and Reasoning-Guided Search in Claim Verification
von: Hu, Qisheng, et al.
Veröffentlicht: (2025)
von: Hu, Qisheng, et al.
Veröffentlicht: (2025)
Enhancing Tabular Anomaly Detection via Pseudo-Label-Guided Generation
von: Huang, Wei, et al.
Veröffentlicht: (2026)
von: Huang, Wei, et al.
Veröffentlicht: (2026)
NMR-Solver: Automated Structure Elucidation via Large-Scale Spectral Matching and Physics-Guided Fragment Optimization
von: Jin, Yongqi, et al.
Veröffentlicht: (2025)
von: Jin, Yongqi, et al.
Veröffentlicht: (2025)
CuSearch: Curriculum Rollout Sampling via Search Depth for Agentic RAG
von: Shen, Jianghan, et al.
Veröffentlicht: (2026)
von: Shen, Jianghan, et al.
Veröffentlicht: (2026)
BEAR: Towards Beam-Search-Aware Optimization for Recommendation with Large Language Models
von: Yang, Weiqin, et al.
Veröffentlicht: (2026)
von: Yang, Weiqin, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Breaking the Curse of Repulsion: Optimistic Distributionally Robust Policy Optimization for Off-Policy Generative Recommendation
von: Jiang, Jie, et al.
Veröffentlicht: (2026) -
S-GRec: Personalized Semantic-Aware Generative Recommendation with Asymmetric Advantage
von: Jiang, Jie, et al.
Veröffentlicht: (2026) -
Spend Less, Reason Better: Budget-Aware Value Tree Search for LLM Agents
von: Li, Yushu, et al.
Veröffentlicht: (2026) -
Adaptive Discovering and Merging for Incremental Novel Class Discovery
von: Chen, Guangyao, et al.
Veröffentlicht: (2024) -
Large Language Model for Verilog Generation with Code-Structure-Guided Reinforcement Learning
von: Wang, Ning, et al.
Veröffentlicht: (2024)