SpecSteer: Synergizing Local Context and Global Reasoning for Efficient Personalized Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Lv, Hang, Liang, Sheng, Wang, Hao, Zhang, Yongyue, Gu, Hongchao, Guo, Wei, Lian, Defu, Liu, Yong, Chen, Enhong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
CoSteer: Collaborative Decoding-Time Personalization via Local Delta Steering
por: Lv, Hang, et al.
Publicado: (2025)
por: Lv, Hang, et al.
Publicado: (2025)
IE as Cache: Information Extraction Enhanced Agentic Reasoning
por: Lv, Hang, et al.
Publicado: (2026)
por: Lv, Hang, et al.
Publicado: (2026)
RAPID: Efficient Retrieval-Augmented Long Text Generation with Writing Planning and Information Discovery
por: Gu, Hongchao, et al.
Publicado: (2025)
por: Gu, Hongchao, et al.
Publicado: (2025)
Learning from Emptiness: De-biasing Listwise Rerankers with Content-Agnostic Probability Calibration
por: Lv, Hang, et al.
Publicado: (2026)
por: Lv, Hang, et al.
Publicado: (2026)
Adaptive Schema-aware Event Extraction with Retrieval-Augmented Generation
por: Liang, Sheng, et al.
Publicado: (2025)
por: Liang, Sheng, et al.
Publicado: (2025)
Efficient Machine Unlearning via Influence Approximation
por: Liu, Jiawei, et al.
Publicado: (2025)
por: Liu, Jiawei, et al.
Publicado: (2025)
Is Softmax Loss All You Need? A Principled Analysis of Softmax-family Loss
por: Pu, Yuanhao, et al.
Publicado: (2026)
por: Pu, Yuanhao, et al.
Publicado: (2026)
Beyond Surrogates: A Quantitative Analysis for Inter-Metric Relationships
por: Pu, Yuanhao, et al.
Publicado: (2026)
por: Pu, Yuanhao, et al.
Publicado: (2026)
SPARD: Self-Paced Curriculum for RL Alignment via Integrating Reward Dynamics and Data Utility
por: Zhi, Xuyang, et al.
Publicado: (2026)
por: Zhi, Xuyang, et al.
Publicado: (2026)
DLF: Enhancing Explicit-Implicit Interaction via Dynamic Low-Order-Aware Fusion for CTR Prediction
por: Wang, Kefan, et al.
Publicado: (2025)
por: Wang, Kefan, et al.
Publicado: (2025)
Efficient Personalized Reranking with Semi-Autoregressive Generation and Online Knowledge Distillation
por: Cheng, Kai, et al.
Publicado: (2026)
por: Cheng, Kai, et al.
Publicado: (2026)
END4Rec: Efficient Noise-Decoupling for Multi-Behavior Sequential Recommendation
por: Han, Yongqiang, et al.
Publicado: (2024)
por: Han, Yongqiang, et al.
Publicado: (2024)
Denoising Pre-Training and Customized Prompt Learning for Efficient Multi-Behavior Sequential Recommendation
por: Wang, Hao, et al.
Publicado: (2024)
por: Wang, Hao, et al.
Publicado: (2024)
Learning to Substitute Components for Compositional Generalization
por: Li, Zhaoyi, et al.
Publicado: (2025)
por: Li, Zhaoyi, et al.
Publicado: (2025)
Dataset Regeneration for Sequential Recommendation
por: Yin, Mingjia, et al.
Publicado: (2024)
por: Yin, Mingjia, et al.
Publicado: (2024)
FuXi-β: Towards a Lightweight and Fast Large-Scale Generative Recommendation Model
por: Ye, Yufei, et al.
Publicado: (2025)
por: Ye, Yufei, et al.
Publicado: (2025)
LLM Cache Bandit Revisited: Addressing Query Heterogeneity for Cost-Effective LLM Inference
por: Yang, Hantao, et al.
Publicado: (2025)
por: Yang, Hantao, et al.
Publicado: (2025)
Securing Recommender System via Cooperative Training
por: Wang, Qingyang, et al.
Publicado: (2024)
por: Wang, Qingyang, et al.
Publicado: (2024)
Generative Large Recommendation Models: Emerging Trends in LLMs for Recommendation
por: Wang, Hao, et al.
Publicado: (2025)
por: Wang, Hao, et al.
Publicado: (2025)
Killing Two Birds with One Stone: Unifying Retrieval and Ranking with a Single Generative Recommendation Model
por: Zhang, Luankang, et al.
Publicado: (2025)
por: Zhang, Luankang, et al.
Publicado: (2025)
Effective and Efficient Schema-aware Information Extraction Using On-Device Large Language Models
por: Wen, Zhihao, et al.
Publicado: (2025)
por: Wen, Zhihao, et al.
Publicado: (2025)
Query-Centric Graph Retrieval Augmented Generation
por: Wu, Yaxiong, et al.
Publicado: (2025)
por: Wu, Yaxiong, et al.
Publicado: (2025)
MLDocRAG: Multimodal Long-Context Document Retrieval Augmented Generation
por: Zhang, Yongyue, et al.
Publicado: (2026)
por: Zhang, Yongyue, et al.
Publicado: (2026)
Learning Deep Tree-based Retriever for Efficient Recommendation: Theory and Method
por: Liu, Ze, et al.
Publicado: (2024)
por: Liu, Ze, et al.
Publicado: (2024)
A Universal Framework for Compressing Embeddings in CTR Prediction
por: Wang, Kefan, et al.
Publicado: (2025)
por: Wang, Kefan, et al.
Publicado: (2025)
Multi-granularity Interest Retrieval and Refinement Network for Long-Term User Behavior Modeling in CTR Prediction
por: Xu, Xiang, et al.
Publicado: (2024)
por: Xu, Xiang, et al.
Publicado: (2024)
Learning Partially Aligned Item Representation for Cross-Domain Sequential Recommendation
por: Yin, Mingjia, et al.
Publicado: (2024)
por: Yin, Mingjia, et al.
Publicado: (2024)
SGMem: Sentence Graph Memory for Long-Term Conversational Agents
por: Wu, Yaxiong, et al.
Publicado: (2025)
por: Wu, Yaxiong, et al.
Publicado: (2025)
PiXTime: A Model for Federated Time Series Forecasting with Heterogeneous Data across Nodes
por: Zhou, Yiming, et al.
Publicado: (2026)
por: Zhou, Yiming, et al.
Publicado: (2026)
Breaking Determinism: Fuzzy Modeling of Sequential Recommendation Using Discrete State Space Diffusion Model
por: Xie, Wenjia, et al.
Publicado: (2024)
por: Xie, Wenjia, et al.
Publicado: (2024)
MDAP: A Multi-view Disentangled and Adaptive Preference Learning Framework for Cross-Domain Recommendation
por: Tong, Junxiong, et al.
Publicado: (2024)
por: Tong, Junxiong, et al.
Publicado: (2024)
From Feature Interaction to Feature Generation: A Generative Paradigm of CTR Prediction Models
por: Yin, Mingjia, et al.
Publicado: (2025)
por: Yin, Mingjia, et al.
Publicado: (2025)
Analytical and Empirical Study of Herding Effects in Recommendation Systems
por: Xie, Hong, et al.
Publicado: (2024)
por: Xie, Hong, et al.
Publicado: (2024)
Multi-agent Multi-armed Bandits with Stochastic Sharable Arm Capacities
por: Xie, Hong, et al.
Publicado: (2024)
por: Xie, Hong, et al.
Publicado: (2024)
Synergizing Systemic Inflammation and Multimodal Ultrasound: Methodological Insights Into Predicting Carotid Plaque Vulnerability
por: Hongchao Lu, et al.
Publicado: (2026)
por: Hongchao Lu, et al.
Publicado: (2026)
Model Specific Task Similarity for Vision Language Model Selection via Layer Conductance
por: Yang, Wei, et al.
Publicado: (2026)
por: Yang, Wei, et al.
Publicado: (2026)
Optimizing Sequential Recommendation Models with Scaling Laws and Approximate Entropy
por: Shen, Tingjia, et al.
Publicado: (2024)
por: Shen, Tingjia, et al.
Publicado: (2024)
Prompting is not Enough: Exploring Knowledge Integration and Controllable Generation
por: Shen, Tingjia, et al.
Publicado: (2025)
por: Shen, Tingjia, et al.
Publicado: (2025)
FuXi-Linear: Unleashing the Power of Linear Attention in Long-term Time-aware Sequential Recommendation
por: Ye, Yufei, et al.
Publicado: (2026)
por: Ye, Yufei, et al.
Publicado: (2026)
Entropy Law: The Story Behind Data Compression and LLM Performance
por: Yin, Mingjia, et al.
Publicado: (2024)
por: Yin, Mingjia, et al.
Publicado: (2024)
Ejemplares similares
-
CoSteer: Collaborative Decoding-Time Personalization via Local Delta Steering
por: Lv, Hang, et al.
Publicado: (2025) -
IE as Cache: Information Extraction Enhanced Agentic Reasoning
por: Lv, Hang, et al.
Publicado: (2026) -
RAPID: Efficient Retrieval-Augmented Long Text Generation with Writing Planning and Information Discovery
por: Gu, Hongchao, et al.
Publicado: (2025) -
Learning from Emptiness: De-biasing Listwise Rerankers with Content-Agnostic Probability Calibration
por: Lv, Hang, et al.
Publicado: (2026) -
Adaptive Schema-aware Event Extraction with Retrieval-Augmented Generation
por: Liang, Sheng, et al.
Publicado: (2025)