Salvato in:
| Autori principali: | Xie, Encheng, Sun, Yihang, Feng, Tao, You, Jiaxuan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2511.08590 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Router-R1: Teaching LLMs Multi-Round Routing and Aggregation via Reinforcement Learning
di: Zhang, Haozhen, et al.
Pubblicazione: (2025)
di: Zhang, Haozhen, et al.
Pubblicazione: (2025)
PersonalizedRouter: Personalized LLM Routing via Graph-based User Preference Modeling
di: Dai, Zhongjie, et al.
Pubblicazione: (2025)
di: Dai, Zhongjie, et al.
Pubblicazione: (2025)
GraphEval: A Lightweight Graph-Based LLM Framework for Idea Evaluation
di: Feng, Tao, et al.
Pubblicazione: (2025)
di: Feng, Tao, et al.
Pubblicazione: (2025)
AcademicEval: Live Long-Context LLM Benchmark
di: Zhang, Haozhen, et al.
Pubblicazione: (2025)
di: Zhang, Haozhen, et al.
Pubblicazione: (2025)
Probing the Knowledge Boundary: An Interactive Agentic Framework for Deep Knowledge Extraction
di: Yang, Yuheng, et al.
Pubblicazione: (2026)
di: Yang, Yuheng, et al.
Pubblicazione: (2026)
Find Your Optimal Teacher: Personalized Data Synthesis via Router-Guided Multi-Teacher Distillation
di: Zhang, Hengyuan, et al.
Pubblicazione: (2025)
di: Zhang, Hengyuan, et al.
Pubblicazione: (2025)
Graph of Records: Boosting Retrieval Augmented Generation for Long-context Summarization with Graphs
di: Zhang, Haozhen, et al.
Pubblicazione: (2024)
di: Zhang, Haozhen, et al.
Pubblicazione: (2024)
ResearchArcade: Graph Interface for Academic Tasks
di: Xu, Jingjun, et al.
Pubblicazione: (2025)
di: Xu, Jingjun, et al.
Pubblicazione: (2025)
LLM Router: Rethinking Routing with Prefill Activations
di: Varshney, Tanay, et al.
Pubblicazione: (2026)
di: Varshney, Tanay, et al.
Pubblicazione: (2026)
OrcaRouter: A Production-Oriented LLM Router with Hybrid Offline-Online Learning
di: Bao, Zhenghua, et al.
Pubblicazione: (2026)
di: Bao, Zhenghua, et al.
Pubblicazione: (2026)
On the Way to LLM Personalization: Learning to Remember User Conversations
di: Magister, Lucie Charlotte, et al.
Pubblicazione: (2024)
di: Magister, Lucie Charlotte, et al.
Pubblicazione: (2024)
Scaling LLM Multi-turn RL with End-to-end Summarization-based Context Management
di: Lu, Miao, et al.
Pubblicazione: (2025)
di: Lu, Miao, et al.
Pubblicazione: (2025)
MINT: Evaluating LLMs in Multi-turn Interaction with Tools and Language Feedback
di: Wang, Xingyao, et al.
Pubblicazione: (2023)
di: Wang, Xingyao, et al.
Pubblicazione: (2023)
AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
di: Ma, Chang, et al.
Pubblicazione: (2024)
di: Ma, Chang, et al.
Pubblicazione: (2024)
ResearchTown: Simulator of Human Research Community
di: Yu, Haofei, et al.
Pubblicazione: (2024)
di: Yu, Haofei, et al.
Pubblicazione: (2024)
Multi-turn Training with Basic Human Feedback Helps Little on LLM Reasoning
di: Liu, Qiang, et al.
Pubblicazione: (2025)
di: Liu, Qiang, et al.
Pubblicazione: (2025)
Skill-SD: Skill-Conditioned Self-Distillation for Multi-turn LLM Agents
di: Wang, Hao, et al.
Pubblicazione: (2026)
di: Wang, Hao, et al.
Pubblicazione: (2026)
DFA-RAG: Conversational Semantic Router for Large Language Model with Definite Finite Automaton
di: Sun, Yiyou, et al.
Pubblicazione: (2024)
di: Sun, Yiyou, et al.
Pubblicazione: (2024)
Stop Listening to Me! How Multi-turn Conversations Can Degrade LLM Reliability
di: Guo, Kevin H., et al.
Pubblicazione: (2026)
di: Guo, Kevin H., et al.
Pubblicazione: (2026)
User-LLM: Efficient LLM Contextualization with User Embeddings
di: Ning, Lin, et al.
Pubblicazione: (2024)
di: Ning, Lin, et al.
Pubblicazione: (2024)
Stabilizing Long-term Multi-turn Reinforcement Learning with Gated Rewards
di: Sun, Zetian, et al.
Pubblicazione: (2025)
di: Sun, Zetian, et al.
Pubblicazione: (2025)
RouterDC: Query-Based Router by Dual Contrastive Learning for Assembling Large Language Models
di: Chen, Shuhao, et al.
Pubblicazione: (2024)
di: Chen, Shuhao, et al.
Pubblicazione: (2024)
Glider: Global and Local Instruction-Driven Expert Router
di: Li, Pingzhi, et al.
Pubblicazione: (2024)
di: Li, Pingzhi, et al.
Pubblicazione: (2024)
Learning to Clarify: Multi-turn Conversations with Action-Based Contrastive Self-Training
di: Chen, Maximillian, et al.
Pubblicazione: (2024)
di: Chen, Maximillian, et al.
Pubblicazione: (2024)
Regressing the Relative Future: Efficient Policy Optimization for Multi-turn RLHF
di: Gao, Zhaolin, et al.
Pubblicazione: (2024)
di: Gao, Zhaolin, et al.
Pubblicazione: (2024)
VL-RouterBench: A Benchmark for Vision-Language Model Routing
di: Huang, Zhehao, et al.
Pubblicazione: (2025)
di: Huang, Zhehao, et al.
Pubblicazione: (2025)
Coupling Experts and Routers in Mixture-of-Experts via an Auxiliary Loss
di: Lv, Ang, et al.
Pubblicazione: (2025)
di: Lv, Ang, et al.
Pubblicazione: (2025)
ROMER: Expert Replacement and Router Calibration for Robust MoE LLMs on Analog Compute-in-Memory Systems
di: Zhou, Wenyong, et al.
Pubblicazione: (2026)
di: Zhou, Wenyong, et al.
Pubblicazione: (2026)
How Far Are We From AGI: Are LLMs All We Need?
di: Feng, Tao, et al.
Pubblicazione: (2024)
di: Feng, Tao, et al.
Pubblicazione: (2024)
Routers Learn the Geometry of Their Experts: Geometric Coupling in Sparse Mixture-of-Experts
di: Ahrac, Sagi, et al.
Pubblicazione: (2026)
di: Ahrac, Sagi, et al.
Pubblicazione: (2026)
RouteNator: A Router-Based Multi-Modal Architecture for Generating Synthetic Training Data for Function Calling LLMs
di: Belavadi, Vibha, et al.
Pubblicazione: (2025)
di: Belavadi, Vibha, et al.
Pubblicazione: (2025)
Cost-Efficient Large Language Model Serving for Multi-turn Conversations with CachedAttention
di: Gao, Bin, et al.
Pubblicazione: (2024)
di: Gao, Bin, et al.
Pubblicazione: (2024)
Turn Waste into Worth: Rectifying Top-$k$ Router of MoE
di: Zeng, Zhiyuan, et al.
Pubblicazione: (2024)
di: Zeng, Zhiyuan, et al.
Pubblicazione: (2024)
MoE-Pruner: Pruning Mixture-of-Experts Large Language Model using the Hints from Its Router
di: Xie, Yanyue, et al.
Pubblicazione: (2024)
di: Xie, Yanyue, et al.
Pubblicazione: (2024)
A Practitioner's Guide to Multi-turn Agentic Reinforcement Learning
di: Wang, Ruiyi, et al.
Pubblicazione: (2025)
di: Wang, Ruiyi, et al.
Pubblicazione: (2025)
Predict, Don't React: Value-Based Safety Forecasting for LLM Streaming
di: Kavumba, Pride, et al.
Pubblicazione: (2026)
di: Kavumba, Pride, et al.
Pubblicazione: (2026)
Paper Copilot: A Self-Evolving and Efficient LLM System for Personalized Academic Assistance
di: Lin, Guanyu, et al.
Pubblicazione: (2024)
di: Lin, Guanyu, et al.
Pubblicazione: (2024)
HISR: Hindsight Information Modulated Segmental Process Rewards For Multi-turn Agentic Reinforcement Learning
di: Lu, Zhicong, et al.
Pubblicazione: (2026)
di: Lu, Zhicong, et al.
Pubblicazione: (2026)
Mitigating Lost in Multi-turn Conversation via Curriculum RL with Verifiable Accuracy and Abstention Rewards
di: Li, Ming, et al.
Pubblicazione: (2025)
di: Li, Ming, et al.
Pubblicazione: (2025)
Debugging Tabular Log as Dynamic Graphs
di: Liang, Chumeng, et al.
Pubblicazione: (2025)
di: Liang, Chumeng, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Router-R1: Teaching LLMs Multi-Round Routing and Aggregation via Reinforcement Learning
di: Zhang, Haozhen, et al.
Pubblicazione: (2025) -
PersonalizedRouter: Personalized LLM Routing via Graph-based User Preference Modeling
di: Dai, Zhongjie, et al.
Pubblicazione: (2025) -
GraphEval: A Lightweight Graph-Based LLM Framework for Idea Evaluation
di: Feng, Tao, et al.
Pubblicazione: (2025) -
AcademicEval: Live Long-Context LLM Benchmark
di: Zhang, Haozhen, et al.
Pubblicazione: (2025) -
Probing the Knowledge Boundary: An Interactive Agentic Framework for Deep Knowledge Extraction
di: Yang, Yuheng, et al.
Pubblicazione: (2026)