Learning to Extract Context for Context-Aware LLM Inference
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Minseon, Caccia, Lucas, Shi, Zhengyan, Pereira, Matheus, Côté, Marc-Alexandre, Yuan, Xingdi, Sordoni, Alessandro |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning to Solve Complex Problems via Dataset Decomposition
by: Zhao, Wanru, et al.
Published: (2026)
by: Zhao, Wanru, et al.
Published: (2026)
BugPilot: Complex Bug Generation for Efficient Learning of SWE Skills
by: Sonwane, Atharv, et al.
Published: (2025)
by: Sonwane, Atharv, et al.
Published: (2025)
Training Plug-n-Play Knowledge Modules with Deep Context Distillation
by: Caccia, Lucas, et al.
Published: (2025)
by: Caccia, Lucas, et al.
Published: (2025)
Guiding Language Model Reasoning with Planning Tokens
by: Wang, Xinyi, et al.
Published: (2023)
by: Wang, Xinyi, et al.
Published: (2023)
Gistify! Codebase-Level Understanding via Runtime Execution
by: Lee, Hyunji, et al.
Published: (2025)
by: Lee, Hyunji, et al.
Published: (2025)
Trade-offs in Ensembling, Merging and Routing Among Parameter-Efficient Experts
by: Lotfi, Sanae, et al.
Published: (2026)
by: Lotfi, Sanae, et al.
Published: (2026)
Exploring Sparse Adapters for Scalable Merging of Parameter Efficient Experts
by: Arnob, Samin Yeasar, et al.
Published: (2025)
by: Arnob, Samin Yeasar, et al.
Published: (2025)
debug-gym: A Text-Based Environment for Interactive Debugging
by: Yuan, Xingdi, et al.
Published: (2025)
by: Yuan, Xingdi, et al.
Published: (2025)
Improving Context-Aware Preference Modeling for Language Models
by: Pitis, Silviu, et al.
Published: (2024)
by: Pitis, Silviu, et al.
Published: (2024)
Policy Improvement using Language Feedback Models
by: Zhong, Victor, et al.
Published: (2024)
by: Zhong, Victor, et al.
Published: (2024)
Test-Time Learning with an Evolving Library
by: Xu, Weijia, et al.
Published: (2026)
by: Xu, Weijia, et al.
Published: (2026)
Dialectics of Alignment: Harnessing Unsafe Knowledge for Dynamic Safety Routing
by: Hashemzadeh, Maryam, et al.
Published: (2026)
by: Hashemzadeh, Maryam, et al.
Published: (2026)
Towards Modular LLMs by Building and Reusing a Library of LoRAs
by: Ostapenko, Oleksiy, et al.
Published: (2024)
by: Ostapenko, Oleksiy, et al.
Published: (2024)
Language-guided Skill Learning with Temporal Variational Inference
by: Fu, Haotian, et al.
Published: (2024)
by: Fu, Haotian, et al.
Published: (2024)
V-STaR: Training Verifiers for Self-Taught Reasoners
by: Hosseini, Arian, et al.
Published: (2024)
by: Hosseini, Arian, et al.
Published: (2024)
Sustainable LLM Inference using Context-Aware Model Switching
by: Yuvarani, et al.
Published: (2026)
by: Yuvarani, et al.
Published: (2026)
GreenServ: Energy-Efficient Context-Aware Dynamic Routing for Multi-Model LLM Inference
by: Ziller, Thomas, et al.
Published: (2026)
by: Ziller, Thomas, et al.
Published: (2026)
A Survey on Model MoErging: Recycling and Routing Among Specialized Experts for Collaborative Learning
by: Yadav, Prateek, et al.
Published: (2024)
by: Yadav, Prateek, et al.
Published: (2024)
Not All LLM Reasoners Are Created Equal
by: Hosseini, Arian, et al.
Published: (2024)
by: Hosseini, Arian, et al.
Published: (2024)
Quest: Query-Aware Sparsity for Efficient Long-Context LLM Inference
by: Tang, Jiaming, et al.
Published: (2024)
by: Tang, Jiaming, et al.
Published: (2024)
InfLLM: Training-Free Long-Context Extrapolation for LLMs with an Efficient Context Memory
by: Xiao, Chaojun, et al.
Published: (2024)
by: Xiao, Chaojun, et al.
Published: (2024)
Prompt Compression with Context-Aware Sentence Encoding for Fast and Improved LLM Inference
by: Liskavets, Barys, et al.
Published: (2024)
by: Liskavets, Barys, et al.
Published: (2024)
ContextFlow: Context-Aware Flow Matching For Trajectory Inference From Spatial Omics Data
by: Rathod, Santanu Subhash, et al.
Published: (2025)
by: Rathod, Santanu Subhash, et al.
Published: (2025)
LLMSteer: Improving Long-Context LLM Inference by Steering Attention on Reused Contexts
by: Gu, Zhuohan, et al.
Published: (2024)
by: Gu, Zhuohan, et al.
Published: (2024)
ContextPilot: Fast Long-Context Inference via Context Reuse
by: Jiang, Yinsicheng, et al.
Published: (2025)
by: Jiang, Yinsicheng, et al.
Published: (2025)
UT-ACA: Uncertainty-Triggered Adaptive Context Allocation for Long-Context Inference
by: Zhou, Lang, et al.
Published: (2026)
by: Zhou, Lang, et al.
Published: (2026)
KVQuant: Towards 10 Million Context Length LLM Inference with KV Cache Quantization
by: Hooper, Coleman, et al.
Published: (2024)
by: Hooper, Coleman, et al.
Published: (2024)
Glance for Context: Learning When to Leverage LLMs for Node-Aware GNN-LLM Fusion
by: Loveland, Donald, et al.
Published: (2025)
by: Loveland, Donald, et al.
Published: (2025)
Context-Aware Trajectory Anomaly Detection
by: Hu, Haoji, et al.
Published: (2024)
by: Hu, Haoji, et al.
Published: (2024)
Putting the Value Back in RL: Better Test-Time Scaling by Unifying LLM Reasoners With Verifiers
by: Sareen, Kusha, et al.
Published: (2025)
by: Sareen, Kusha, et al.
Published: (2025)
HGCA: Hybrid GPU-CPU Attention for Long Context LLM Inference
by: Deng, Weishu, et al.
Published: (2025)
by: Deng, Weishu, et al.
Published: (2025)
AccLLM: Accelerating Long-Context LLM Inference Via Algorithm-Hardware Co-Design
by: Liang, Yanbiao, et al.
Published: (2025)
by: Liang, Yanbiao, et al.
Published: (2025)
Context-Aware Deep Lagrangian Networks for Model Predictive Control
by: Schulze, Lucas, et al.
Published: (2025)
by: Schulze, Lucas, et al.
Published: (2025)
Protein Representation Learning by Capturing Protein Sequence-Structure-Function Relationship
by: Ko, Eunji, et al.
Published: (2024)
by: Ko, Eunji, et al.
Published: (2024)
Towards Foundation Inference Models that Learn ODEs In-Context
by: Mauel, Maximilian, et al.
Published: (2025)
by: Mauel, Maximilian, et al.
Published: (2025)
Can Transformers Learn Full Bayesian Inference in Context?
by: Reuter, Arik, et al.
Published: (2025)
by: Reuter, Arik, et al.
Published: (2025)
OPEx: A Component-Wise Analysis of LLM-Centric Agents in Embodied Instruction Following
by: Shi, Haochen, et al.
Published: (2024)
by: Shi, Haochen, et al.
Published: (2024)
Megalodon: Efficient LLM Pretraining and Inference with Unlimited Context Length
by: Ma, Xuezhe, et al.
Published: (2024)
by: Ma, Xuezhe, et al.
Published: (2024)
CurveRL: Principled Distribution-Aware Context Reweighting for LLM Reasoning
by: Sun, Ke, et al.
Published: (2026)
by: Sun, Ke, et al.
Published: (2026)
Enhancing Hyperedge Prediction with Context-Aware Self-Supervised Learning
by: Ko, Yunyong, et al.
Published: (2023)
by: Ko, Yunyong, et al.
Published: (2023)
Similar Items
-
Learning to Solve Complex Problems via Dataset Decomposition
by: Zhao, Wanru, et al.
Published: (2026) -
BugPilot: Complex Bug Generation for Efficient Learning of SWE Skills
by: Sonwane, Atharv, et al.
Published: (2025) -
Training Plug-n-Play Knowledge Modules with Deep Context Distillation
by: Caccia, Lucas, et al.
Published: (2025) -
Guiding Language Model Reasoning with Planning Tokens
by: Wang, Xinyi, et al.
Published: (2023) -
Gistify! Codebase-Level Understanding via Runtime Execution
by: Lee, Hyunji, et al.
Published: (2025)