CONCUR: A Framework for Continual Constrained and Unconstrained Routing
Fuente:
arXiv
Salvato in:
| Autori principali: | Chen, Peter Baile, Li, Weiyue, Roth, Dan, Cafarella, Michael, Madden, Samuel, Andreas, Jacob |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Log-Augmented Generation: Scaling Test-Time Reasoning with Reusable Computation
di: Chen, Peter Baile, et al.
Pubblicazione: (2025)
di: Chen, Peter Baile, et al.
Pubblicazione: (2025)
Can we Retrieve Everything All at Once? ARM: An Alignment-Oriented LLM-based Retrieval Method
di: Chen, Peter Baile, et al.
Pubblicazione: (2025)
di: Chen, Peter Baile, et al.
Pubblicazione: (2025)
EnrichIndex: Using LLMs to Enrich Retrieval Indices Offline
di: Chen, Peter Baile, et al.
Pubblicazione: (2025)
di: Chen, Peter Baile, et al.
Pubblicazione: (2025)
BEAVER: An Enterprise Benchmark for Text-to-SQL
di: Chen, Peter Baile, et al.
Pubblicazione: (2024)
di: Chen, Peter Baile, et al.
Pubblicazione: (2024)
PEEK: Context Map as an Orientation Cache for Long-Context LLM Agents
di: Gu, Zhuohan, et al.
Pubblicazione: (2026)
di: Gu, Zhuohan, et al.
Pubblicazione: (2026)
MDCR: A Dataset for Multi-Document Conditional Reasoning
di: Chen, Peter Baile, et al.
Pubblicazione: (2024)
di: Chen, Peter Baile, et al.
Pubblicazione: (2024)
Is Table Retrieval a Solved Problem? Exploring Join-Aware Multi-Table Retrieval
di: Chen, Peter Baile, et al.
Pubblicazione: (2024)
di: Chen, Peter Baile, et al.
Pubblicazione: (2024)
SynapseRoute: An Auto-Route Switching Framework on Dual-State Large Language Model
di: Zhang, Wencheng, et al.
Pubblicazione: (2025)
di: Zhang, Wencheng, et al.
Pubblicazione: (2025)
View From Above: A Framework for Evaluating Distribution Shifts in Model Behavior
di: Chopra, Tanush, et al.
Pubblicazione: (2024)
di: Chopra, Tanush, et al.
Pubblicazione: (2024)
OpenEstimate: Evaluating LLMs on Reasoning Under Uncertainty with Real-World Data
di: Renda, Alana, et al.
Pubblicazione: (2025)
di: Renda, Alana, et al.
Pubblicazione: (2025)
Counterfactual Reasoning with Knowledge Graph Embeddings
di: Zellinger, Lena, et al.
Pubblicazione: (2024)
di: Zellinger, Lena, et al.
Pubblicazione: (2024)
Learning to Route for Dynamic Adapter Composition in Continual Learning with Language Models
di: Araujo, Vladimir, et al.
Pubblicazione: (2024)
di: Araujo, Vladimir, et al.
Pubblicazione: (2024)
Evaluating NL2SQL via SQL2NL
di: Safarzadeh, Mohammadtaher, et al.
Pubblicazione: (2025)
di: Safarzadeh, Mohammadtaher, et al.
Pubblicazione: (2025)
Toward In-Context Teaching: Adapting Examples to Students' Misconceptions
di: Ross, Alexis, et al.
Pubblicazione: (2024)
di: Ross, Alexis, et al.
Pubblicazione: (2024)
Algorithmic Capabilities of Random Transformers
di: Zhong, Ziqian, et al.
Pubblicazione: (2024)
di: Zhong, Ziqian, et al.
Pubblicazione: (2024)
ReEval: Automatic Hallucination Evaluation for Retrieval-Augmented Large Language Models via Transferable Adversarial Attacks
di: Yu, Xiaodong, et al.
Pubblicazione: (2023)
di: Yu, Xiaodong, et al.
Pubblicazione: (2023)
Training Language Models to Explain Their Own Computations
di: Li, Belinda Z., et al.
Pubblicazione: (2025)
di: Li, Belinda Z., et al.
Pubblicazione: (2025)
Analytic Subspace Routing: How Recursive Least Squares Works in Continual Learning of Large Language Model
di: Tong, Kai, et al.
Pubblicazione: (2025)
di: Tong, Kai, et al.
Pubblicazione: (2025)
(How) Do Language Models Track State?
di: Li, Belinda Z., et al.
Pubblicazione: (2025)
di: Li, Belinda Z., et al.
Pubblicazione: (2025)
Enhancing Temporal Understanding in LLMs for Semi-structured Tables
di: Deng, Irwin, et al.
Pubblicazione: (2024)
di: Deng, Irwin, et al.
Pubblicazione: (2024)
RouteLLM: Learning to Route LLMs with Preference Data
di: Ong, Isaac, et al.
Pubblicazione: (2024)
di: Ong, Isaac, et al.
Pubblicazione: (2024)
Sample-Efficient Online Learning in LM Agents via Hindsight Trajectory Rewriting
di: Hu, Michael Y., et al.
Pubblicazione: (2025)
di: Hu, Michael Y., et al.
Pubblicazione: (2025)
A Hitchhiker's Guide to Scaling Law Estimation
di: Choshen, Leshem, et al.
Pubblicazione: (2024)
di: Choshen, Leshem, et al.
Pubblicazione: (2024)
Lexicon-Level Contrastive Visual-Grounding Improves Language Modeling
di: Zhuang, Chengxu, et al.
Pubblicazione: (2024)
di: Zhuang, Chengxu, et al.
Pubblicazione: (2024)
Tree-based Dialogue Reinforced Policy Optimization for Red-Teaming Attacks
di: Guo, Ruohao, et al.
Pubblicazione: (2025)
di: Guo, Ruohao, et al.
Pubblicazione: (2025)
Explaining Datasets in Words: Statistical Models with Natural Language Parameters
di: Zhong, Ruiqi, et al.
Pubblicazione: (2024)
di: Zhong, Ruiqi, et al.
Pubblicazione: (2024)
Discovering Latent Knowledge in Language Models Without Supervision
di: Burns, Collin, et al.
Pubblicazione: (2022)
di: Burns, Collin, et al.
Pubblicazione: (2022)
Constrained Entropic Unlearning: A Primal-Dual Framework for Large Language Models
di: Entesari, Taha, et al.
Pubblicazione: (2025)
di: Entesari, Taha, et al.
Pubblicazione: (2025)
H-STAR: LLM-driven Hybrid SQL-Text Adaptive Reasoning on Tables
di: Abhyankar, Nikhil, et al.
Pubblicazione: (2024)
di: Abhyankar, Nikhil, et al.
Pubblicazione: (2024)
Bias Is a Subspace, Not a Coordinate: A Geometric Rethinking of Post-hoc Debiasing in Vision-Language Models
di: Zhao, Dachuan, et al.
Pubblicazione: (2025)
di: Zhao, Dachuan, et al.
Pubblicazione: (2025)
Dynamic Latent Routing
di: Yu, Fangyuan, et al.
Pubblicazione: (2026)
di: Yu, Fangyuan, et al.
Pubblicazione: (2026)
Chain of Simulation: A Dual-Mode Reasoning Framework for Large Language Models with Dynamic Problem Routing
di: Sheikhi, Saeid
Pubblicazione: (2026)
di: Sheikhi, Saeid
Pubblicazione: (2026)
RewardUQ: A Unified Framework for Uncertainty-Aware Reward Models
di: Yang, Daniel, et al.
Pubblicazione: (2026)
di: Yang, Daniel, et al.
Pubblicazione: (2026)
Token-Level LLM Collaboration via FusionRoute
di: Xiong, Nuoya, et al.
Pubblicazione: (2026)
di: Xiong, Nuoya, et al.
Pubblicazione: (2026)
Continual Knowledge Updating in LLM Systems: Learning Through Multi-Timescale Memory Dynamics
di: Pattichis, Andreas, et al.
Pubblicazione: (2026)
di: Pattichis, Andreas, et al.
Pubblicazione: (2026)
Fewer Truncations Improve Language Modeling
di: Ding, Hantian, et al.
Pubblicazione: (2024)
di: Ding, Hantian, et al.
Pubblicazione: (2024)
Merge, Then Compress: Demystify Efficient SMoE with Hints from Its Routing Policy
di: Li, Pingzhi, et al.
Pubblicazione: (2023)
di: Li, Pingzhi, et al.
Pubblicazione: (2023)
Routing-Free Mixture-of-Experts
di: Liu, Yilun, et al.
Pubblicazione: (2026)
di: Liu, Yilun, et al.
Pubblicazione: (2026)
Multilingual Routing in Mixture-of-Experts
di: Bandarkar, Lucas, et al.
Pubblicazione: (2025)
di: Bandarkar, Lucas, et al.
Pubblicazione: (2025)
Advancing MoE Efficiency: A Collaboration-Constrained Routing (C2R) Strategy for Better Expert Parallelism Design
di: Zhang, Mohan, et al.
Pubblicazione: (2025)
di: Zhang, Mohan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Log-Augmented Generation: Scaling Test-Time Reasoning with Reusable Computation
di: Chen, Peter Baile, et al.
Pubblicazione: (2025) -
Can we Retrieve Everything All at Once? ARM: An Alignment-Oriented LLM-based Retrieval Method
di: Chen, Peter Baile, et al.
Pubblicazione: (2025) -
EnrichIndex: Using LLMs to Enrich Retrieval Indices Offline
di: Chen, Peter Baile, et al.
Pubblicazione: (2025) -
BEAVER: An Enterprise Benchmark for Text-to-SQL
di: Chen, Peter Baile, et al.
Pubblicazione: (2024) -
PEEK: Context Map as an Orientation Cache for Long-Context LLM Agents
di: Gu, Zhuohan, et al.
Pubblicazione: (2026)