Quality Over Clicks: Intrinsic Quality-Driven Iterative Reinforcement Learning for Cold-Start E-Commerce Query Suggestion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Qi, Xiao, Kejun, Zhao, Huaipeng, Luo, Tao, Zeng, Xiaoyi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Shopping Companion: Benchmarking and Training LLM Agents for Long-Horizon Preference-Grounded E-Commerce Tasks
von: Yu, Zijian, et al.
Veröffentlicht: (2026)
von: Yu, Zijian, et al.
Veröffentlicht: (2026)
ProductResearch: Training E-Commerce Deep Research Agents via Multi-Agent Synthetic Trajectory Distillation
von: Wang, Jiangyuan, et al.
Veröffentlicht: (2026)
von: Wang, Jiangyuan, et al.
Veröffentlicht: (2026)
ShoppingBench: A Real-World Intent-Grounded Shopping Benchmark for LLM-based Agents
von: Wang, Jiangyuan, et al.
Veröffentlicht: (2025)
von: Wang, Jiangyuan, et al.
Veröffentlicht: (2025)
Advancing Multimodal Reasoning via Reinforcement Learning with Cold Start
von: Wei, Lai, et al.
Veröffentlicht: (2025)
von: Wei, Lai, et al.
Veröffentlicht: (2025)
From Clicks to Preference: A Multi-stage Alignment Framework for Generative Query Suggestion in Conversational System
von: Yin, Junhao, et al.
Veröffentlicht: (2025)
von: Yin, Junhao, et al.
Veröffentlicht: (2025)
Better Datasets Start From RefineLab: Automatic Optimization for High-Quality Dataset Refinement
von: Luo, Xiaonan, et al.
Veröffentlicht: (2025)
von: Luo, Xiaonan, et al.
Veröffentlicht: (2025)
Advancing Multimodal Reasoning: From Optimized Cold Start to Staged Reinforcement Learning
von: Chen, Shuang, et al.
Veröffentlicht: (2025)
von: Chen, Shuang, et al.
Veröffentlicht: (2025)
Identifying High Consideration E-Commerce Search Queries
von: Chen, Zhiyu, et al.
Veröffentlicht: (2024)
von: Chen, Zhiyu, et al.
Veröffentlicht: (2024)
Query Suggestion for Retrieval-Augmented Generation via Dynamic In-Context Learning
von: Spaeh, Fabian, et al.
Veröffentlicht: (2026)
von: Spaeh, Fabian, et al.
Veröffentlicht: (2026)
Context Over Compute Human-in-the-Loop Outperforms Iterative Chain-of-Thought Prompting in Interview Answer Quality
von: Zhu, Kewen, et al.
Veröffentlicht: (2026)
von: Zhu, Kewen, et al.
Veröffentlicht: (2026)
Modeling Data Diversity for Joint Instance and Verbalizer Selection in Cold-Start Scenarios
von: Chakraborty, Mohna, et al.
Veröffentlicht: (2025)
von: Chakraborty, Mohna, et al.
Veröffentlicht: (2025)
Enhancing the Capability and Robustness of Large Language Models through Reinforcement Learning-Driven Query Refinement
von: Wang, Xiaohua, et al.
Veröffentlicht: (2024)
von: Wang, Xiaohua, et al.
Veröffentlicht: (2024)
RouteProfile: Graph-Based Profiling for Cold-Start LLM Routing
von: Xu, Jingjun, et al.
Veröffentlicht: (2026)
von: Xu, Jingjun, et al.
Veröffentlicht: (2026)
CSRM-LLM: Embracing Multilingual LLMs for Cold-Start Relevance Matching in Emerging E-commerce Markets
von: Wang, Yujing, et al.
Veröffentlicht: (2025)
von: Wang, Yujing, et al.
Veröffentlicht: (2025)
RIVAL: Reinforcement Learning with Iterative and Adversarial Optimization for Machine Translation
von: Li, Tianjiao, et al.
Veröffentlicht: (2025)
von: Li, Tianjiao, et al.
Veröffentlicht: (2025)
Examples as the Prompt: A Scalable Approach for Efficient LLM Adaptation in E-Commerce
von: Zeng, Jingying, et al.
Veröffentlicht: (2025)
von: Zeng, Jingying, et al.
Veröffentlicht: (2025)
LLaSA: Large Language and E-Commerce Shopping Assistant
von: Zhang, Shuo, et al.
Veröffentlicht: (2024)
von: Zhang, Shuo, et al.
Veröffentlicht: (2024)
From Zero to Hero: Cold-Start Anomaly Detection
von: Reiss, Tal, et al.
Veröffentlicht: (2024)
von: Reiss, Tal, et al.
Veröffentlicht: (2024)
LocalSUG: City-Preference-Enhanced LLM for Query Suggestion in Local-Life Services
von: Chen, Jinwen, et al.
Veröffentlicht: (2026)
von: Chen, Jinwen, et al.
Veröffentlicht: (2026)
QCQA: Quality and Capacity-aware grouped Query Attention
von: Joshi, Vinay, et al.
Veröffentlicht: (2024)
von: Joshi, Vinay, et al.
Veröffentlicht: (2024)
Metis-SPECS: Decoupling Multimodal Learning via Self-distilled Preference-based Cold Start
von: Chen, Kun, et al.
Veröffentlicht: (2025)
von: Chen, Kun, et al.
Veröffentlicht: (2025)
STENCIL: Submodular Mutual Information Based Weak Supervision for Cold-Start Active Learning
von: Beck, Nathan, et al.
Veröffentlicht: (2024)
von: Beck, Nathan, et al.
Veröffentlicht: (2024)
Curriculum Learning with Quality-Driven Data Selection
von: Wu, Biao, et al.
Veröffentlicht: (2024)
von: Wu, Biao, et al.
Veröffentlicht: (2024)
LLM-Driven E-Commerce Marketing Content Optimization: Balancing Creativity and Conversion
von: Yang, Haowei, et al.
Veröffentlicht: (2025)
von: Yang, Haowei, et al.
Veröffentlicht: (2025)
EPM-RL: Reinforcement Learning for On-Premise Product Mapping in E-Commerce
von: Yu, Minhyeong, et al.
Veröffentlicht: (2026)
von: Yu, Minhyeong, et al.
Veröffentlicht: (2026)
RAGSys: Item-Cold-Start Recommender as RAG System
von: Contal, Emile, et al.
Veröffentlicht: (2024)
von: Contal, Emile, et al.
Veröffentlicht: (2024)
Preserving Multilingual Quality While Tuning Query Encoder on English Only
von: Vasilyev, Oleg, et al.
Veröffentlicht: (2024)
von: Vasilyev, Oleg, et al.
Veröffentlicht: (2024)
QCRD: Quality-guided Contrastive Rationale Distillation for Large Language Models
von: Wang, Wei, et al.
Veröffentlicht: (2024)
von: Wang, Wei, et al.
Veröffentlicht: (2024)
PRIME: Policy-Reinforced Iterative Multi-agent Execution for Algorithmic Reasoning in Large Language Models
von: Xu, Jiawei, et al.
Veröffentlicht: (2026)
von: Xu, Jiawei, et al.
Veröffentlicht: (2026)
When to Trust LLMs: Aligning Confidence with Response Quality
von: Tao, Shuchang, et al.
Veröffentlicht: (2024)
von: Tao, Shuchang, et al.
Veröffentlicht: (2024)
Thinking Broad, Acting Fast: Latent Reasoning Distillation from Multi-Perspective Chain-of-Thought for E-Commerce Relevance
von: Qiu, Baopu, et al.
Veröffentlicht: (2026)
von: Qiu, Baopu, et al.
Veröffentlicht: (2026)
Hybrid Querying Over Relational Databases and Large Language Models
von: Zhao, Fuheng, et al.
Veröffentlicht: (2024)
von: Zhao, Fuheng, et al.
Veröffentlicht: (2024)
No Free Lunch in Active Learning: LLM Embedding Quality Dictates Query Strategy Success
von: Rauch, Lukas, et al.
Veröffentlicht: (2025)
von: Rauch, Lukas, et al.
Veröffentlicht: (2025)
Cold-Start Personalization via Training-Free Priors from Structured World Models
von: Bose, Avinandan, et al.
Veröffentlicht: (2026)
von: Bose, Avinandan, et al.
Veröffentlicht: (2026)
MemOrb: A Plug-and-Play Verbal-Reinforcement Memory Layer for E-Commerce Customer Service
von: Huang, Yizhe, et al.
Veröffentlicht: (2025)
von: Huang, Yizhe, et al.
Veröffentlicht: (2025)
Improving Content Recommendation: Knowledge Graph-Based Semantic Contrastive Learning for Diversity and Cold-Start Users
von: Kim, Yejin, et al.
Veröffentlicht: (2024)
von: Kim, Yejin, et al.
Veröffentlicht: (2024)
Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing
von: Ding, Dujian, et al.
Veröffentlicht: (2024)
von: Ding, Dujian, et al.
Veröffentlicht: (2024)
Towards Cold-Start Drafting and Continual Refining: A Value-Driven Memory Approach with Application to NPU Kernel Synthesis
von: Zheng, Yujie, et al.
Veröffentlicht: (2026)
von: Zheng, Yujie, et al.
Veröffentlicht: (2026)
Curiosity-Driven Reinforcement Learning from Human Feedback
von: Sun, Haoran, et al.
Veröffentlicht: (2025)
von: Sun, Haoran, et al.
Veröffentlicht: (2025)
Beyond Scalar Scores: Reinforcement Learning for Error-Aware Quality Estimation of Machine Translation
von: Sindhujan, Archchana, et al.
Veröffentlicht: (2026)
von: Sindhujan, Archchana, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Shopping Companion: Benchmarking and Training LLM Agents for Long-Horizon Preference-Grounded E-Commerce Tasks
von: Yu, Zijian, et al.
Veröffentlicht: (2026) -
ProductResearch: Training E-Commerce Deep Research Agents via Multi-Agent Synthetic Trajectory Distillation
von: Wang, Jiangyuan, et al.
Veröffentlicht: (2026) -
ShoppingBench: A Real-World Intent-Grounded Shopping Benchmark for LLM-based Agents
von: Wang, Jiangyuan, et al.
Veröffentlicht: (2025) -
Advancing Multimodal Reasoning via Reinforcement Learning with Cold Start
von: Wei, Lai, et al.
Veröffentlicht: (2025) -
From Clicks to Preference: A Multi-stage Alignment Framework for Generative Query Suggestion in Conversational System
von: Yin, Junhao, et al.
Veröffentlicht: (2025)