DyBBT: Dynamic Balance via Bandit-inspired Targeting for Dialog Policy with Cognitive Dual-Systems
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Shuyu, Wei, Yifan, Yuan, Jialuo, Wang, Xinru, Zhu, Yanmin, Li, Bin, Liu, Yujie |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HiCoLoRA: Addressing Context-Prompt Misalignment via Hierarchical Collaborative LoRA for Zero-Shot DST
by: Zhang, Shuyu, et al.
Published: (2025)
by: Zhang, Shuyu, et al.
Published: (2025)
DarwinTOD: LLM-driven Lifelong Self-evolution for Task-oriented Dialog Systems
by: Zhang, Shuyu, et al.
Published: (2026)
by: Zhang, Shuyu, et al.
Published: (2026)
Balancing Accuracy and Efficiency in Multi-Turn Intent Classification for LLM-Powered Dialog Systems in Production
by: Liu, Junhua, et al.
Published: (2024)
by: Liu, Junhua, et al.
Published: (2024)
Causal Deconfounding via Confounder Disentanglement for Dual-Target Cross-Domain Recommendation
by: Zhu, Jiajie, et al.
Published: (2024)
by: Zhu, Jiajie, et al.
Published: (2024)
DyVo: Dynamic Vocabularies for Learned Sparse Retrieval with Entities
by: Nguyen, Thong, et al.
Published: (2024)
by: Nguyen, Thong, et al.
Published: (2024)
DyG-RAG: Dynamic Graph Retrieval-Augmented Generation with Event-Centric Reasoning
by: Sun, Qingyun, et al.
Published: (2025)
by: Sun, Qingyun, et al.
Published: (2025)
Cog-RAG: Cognitive-Inspired Dual-Hypergraph with Theme Alignment Retrieval-Augmented Generation
by: Hu, Hao, et al.
Published: (2025)
by: Hu, Hao, et al.
Published: (2025)
Optimal Baseline Corrections for Off-Policy Contextual Bandits
by: Gupta, Shashank, et al.
Published: (2024)
by: Gupta, Shashank, et al.
Published: (2024)
SaFRO: Satisfaction-Aware Fusion via Dual-Relative Policy Optimization for Short-Video Search
by: Zhou, Renzhe, et al.
Published: (2026)
by: Zhou, Renzhe, et al.
Published: (2026)
Long Dialog Summarization: An Analysis
by: Mullick, Ankan, et al.
Published: (2024)
by: Mullick, Ankan, et al.
Published: (2024)
Balancing Fine-tuning and RAG: A Hybrid Strategy for Dynamic LLM Recommendation Updates
by: Meng, Changping, et al.
Published: (2025)
by: Meng, Changping, et al.
Published: (2025)
Modeling Attrition in Recommender Systems with Departing Bandits
by: Ben-Porat, Omer, et al.
Published: (2022)
by: Ben-Porat, Omer, et al.
Published: (2022)
Meta Clustering of Neural Bandits
by: Ban, Yikun, et al.
Published: (2024)
by: Ban, Yikun, et al.
Published: (2024)
Enhancing New-item Fairness in Dynamic Recommender Systems
by: Guo, Huizhong, et al.
Published: (2025)
by: Guo, Huizhong, et al.
Published: (2025)
DARLR: Dual-Agent Offline Reinforcement Learning for Recommender Systems with Dynamic Reward
by: Zhang, Yi, et al.
Published: (2025)
by: Zhang, Yi, et al.
Published: (2025)
LLM-CoT Enhanced Graph Neural Recommendation with Harmonized Group Policy Optimization
by: Luo, Hailong, et al.
Published: (2025)
by: Luo, Hailong, et al.
Published: (2025)
Optimizing User Profiles via Contextual Bandits for Retrieval-Augmented LLM Personalization
by: Du, Linfeng, et al.
Published: (2026)
by: Du, Linfeng, et al.
Published: (2026)
Boosting the Targeted Transferability of Adversarial Examples via Salient Region & Weighted Feature Drop
by: Xu, Shanjun, et al.
Published: (2024)
by: Xu, Shanjun, et al.
Published: (2024)
Quantum Cognition-Inspired EEG-based Recommendation via Graph Neural Networks
by: Han, Jinkun, et al.
Published: (2025)
by: Han, Jinkun, et al.
Published: (2025)
Task-Oriented Dialog Systems for the Senegalese Wolof Language
by: Mbaye, Derguene, et al.
Published: (2024)
by: Mbaye, Derguene, et al.
Published: (2024)
Apollonion: Profile-centric Dialog Agent
by: Chen, Shangyu, et al.
Published: (2024)
by: Chen, Shangyu, et al.
Published: (2024)
Agentic Entropy-Balanced Policy Optimization
by: Dong, Guanting, et al.
Published: (2025)
by: Dong, Guanting, et al.
Published: (2025)
Preference-Consistent Knowledge Distillation for Recommender System
by: Zhu, Zhangchi, et al.
Published: (2023)
by: Zhu, Zhangchi, et al.
Published: (2023)
DARTS: A Dual-View Attack Framework for Targeted Manipulation in Federated Sequential Recommendation
by: Qin, Qitao, et al.
Published: (2025)
by: Qin, Qitao, et al.
Published: (2025)
Graph-Structured Driven Dual Adaptation for Mitigating Popularity Bias
by: Cai, Miaomiao, et al.
Published: (2025)
by: Cai, Miaomiao, et al.
Published: (2025)
SynDy: Synthetic Dynamic Dataset Generation Framework for Misinformation Tasks
by: Shliselberg, Michael, et al.
Published: (2024)
by: Shliselberg, Michael, et al.
Published: (2024)
Low-Rank Online Dynamic Assortment with Dual Contextual Information
by: Lee, Seong Jin, et al.
Published: (2024)
by: Lee, Seong Jin, et al.
Published: (2024)
Counterfactual Multi-player Bandits for Explainable Recommendation Diversification
by: Zhang, Yansen, et al.
Published: (2025)
by: Zhang, Yansen, et al.
Published: (2025)
The Bandit's Blind Spot: The Critical Role of User State Representation in Recommender Systems
by: Pires, Pedro R., et al.
Published: (2026)
by: Pires, Pedro R., et al.
Published: (2026)
Causal Feature Selection Method for Contextual Multi-Armed Bandits in Recommender System
by: Zhao, Zhenyu, et al.
Published: (2024)
by: Zhao, Zhenyu, et al.
Published: (2024)
Fair Agents: Balancing Multistakeholder Alignment in Multi-Agent Personalization Systems
by: Forster, Andrea, et al.
Published: (2026)
by: Forster, Andrea, et al.
Published: (2026)
Reward Balancing Revisited: Enhancing Offline Reinforcement Learning for Recommender Systems
by: Shu, Wenzheng, et al.
Published: (2025)
by: Shu, Wenzheng, et al.
Published: (2025)
Dynamic User Interest Augmentation via Stream Clustering and Memory Networks in Large-Scale Recommender Systems
by: Liu, Peng, et al.
Published: (2024)
by: Liu, Peng, et al.
Published: (2024)
Content Moderation in TV Search: Balancing Policy Compliance, Relevance, and User Experience
by: Hande, Adeep, et al.
Published: (2025)
by: Hande, Adeep, et al.
Published: (2025)
LLM-Driven Dual-Level Multi-Interest Modeling for Recommendation
by: Wang, Ziyan, et al.
Published: (2025)
by: Wang, Ziyan, et al.
Published: (2025)
Calibrated Recommendations with Contextual Bandits
by: Feijer, Diego, et al.
Published: (2025)
by: Feijer, Diego, et al.
Published: (2025)
ACT: Automated Constraint Targeting for Multi-Objective Recommender Systems
by: Chang, Daryl, et al.
Published: (2025)
by: Chang, Daryl, et al.
Published: (2025)
Uplift Modeling for Target User Attacks on Recommender Systems
by: Wang, Wenjie, et al.
Published: (2024)
by: Wang, Wenjie, et al.
Published: (2024)
A New Query Expansion Approach via Agent-Mediated Dialogic Inquiry
by: Seo, Wonduk, et al.
Published: (2025)
by: Seo, Wonduk, et al.
Published: (2025)
A Parameter Update Balancing Algorithm for Multi-task Ranking Models in Recommendation Systems
by: Yuan, Jun, et al.
Published: (2024)
by: Yuan, Jun, et al.
Published: (2024)
Similar Items
-
HiCoLoRA: Addressing Context-Prompt Misalignment via Hierarchical Collaborative LoRA for Zero-Shot DST
by: Zhang, Shuyu, et al.
Published: (2025) -
DarwinTOD: LLM-driven Lifelong Self-evolution for Task-oriented Dialog Systems
by: Zhang, Shuyu, et al.
Published: (2026) -
Balancing Accuracy and Efficiency in Multi-Turn Intent Classification for LLM-Powered Dialog Systems in Production
by: Liu, Junhua, et al.
Published: (2024) -
Causal Deconfounding via Confounder Disentanglement for Dual-Target Cross-Domain Recommendation
by: Zhu, Jiajie, et al.
Published: (2024) -
DyVo: Dynamic Vocabularies for Learned Sparse Retrieval with Entities
by: Nguyen, Thong, et al.
Published: (2024)