Context Length Alone Hurts LLM Performance Despite Perfect Retrieval
Fuente:
arXiv
Saved in:
| Main Authors: | Du, Yufeng, Tian, Minyang, Ronanki, Srikanth, Rongali, Subendhu, Bodapati, Sravan, Galstyan, Aram, Wells, Azton, Schwartz, Roy, Huerta, Eliu A, Peng, Hao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RoPE Distinguishes Neither Positions Nor Tokens in Long Contexts, Provably
by: Du, Yufeng, et al.
Published: (2026)
by: Du, Yufeng, et al.
Published: (2026)
Compress, Gather, and Recompute: REFORMing Long-Context Processing in Transformers
by: Song, Woomin, et al.
Published: (2025)
by: Song, Woomin, et al.
Published: (2025)
Adaptive Video Understanding Agent: Enhancing efficiency with dynamic frame sampling and feedback-driven reasoning
by: Jeoung, Sullam, et al.
Published: (2024)
by: Jeoung, Sullam, et al.
Published: (2024)
DCTX-Conformer: Dynamic context carry-over for low latency unified streaming and non-streaming Conformer ASR
by: Huybrechts, Goeric, et al.
Published: (2023)
by: Huybrechts, Goeric, et al.
Published: (2023)
SeRA: Self-Reviewing and Alignment of Large Language Models using Implicit Reward Margins
by: Ko, Jongwoo, et al.
Published: (2024)
by: Ko, Jongwoo, et al.
Published: (2024)
LAWCAT: Efficient Distillation from Quadratic to Linear Attention with Convolution across Tokens for Long Context Modeling
by: Liu, Zeyu, et al.
Published: (2025)
by: Liu, Zeyu, et al.
Published: (2025)
Accelerated Test-Time Scaling with Model-Free Speculative Sampling
by: Song, Woomin, et al.
Published: (2025)
by: Song, Woomin, et al.
Published: (2025)
AttenGW: A Lightweight Attention-Based Multi-Detector Gravitational-Wave Detection Pipeline
by: Tiki, Victoria, et al.
Published: (2025)
by: Tiki, Victoria, et al.
Published: (2025)
Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark
by: Huybrechts, Goeric, et al.
Published: (2025)
by: Huybrechts, Goeric, et al.
Published: (2025)
Think Clearly: Improving Reasoning via Redundant Token Pruning
by: Choi, Daewon, et al.
Published: (2025)
by: Choi, Daewon, et al.
Published: (2025)
Facilitating Trustworthy Human-Agent Collaboration in LLM-based Multi-Agent System oriented Software Engineering
by: Ronanki, Krishna
Published: (2025)
by: Ronanki, Krishna
Published: (2025)
KG-LLM-Bench: A Scalable Benchmark for Evaluating LLM Reasoning on Textualized Knowledge Graphs
by: Markowitz, Elan, et al.
Published: (2025)
by: Markowitz, Elan, et al.
Published: (2025)
Sequence modeling of higher-order wave modes of binary black hole mergers
by: Tiki, Victoria, et al.
Published: (2024)
by: Tiki, Victoria, et al.
Published: (2024)
Mamba Drafters for Speculative Decoding
by: Choi, Daewon, et al.
Published: (2025)
by: Choi, Daewon, et al.
Published: (2025)
Speech Retrieval-Augmented Generation without Automatic Speech Recognition
by: Min, Do June, et al.
Published: (2024)
by: Min, Do June, et al.
Published: (2024)
When Correct Demonstrations Hurt: Rethinking the Role of Exemplars in In-Context Learning
by: Qiu, Chenghao, et al.
Published: (2026)
by: Qiu, Chenghao, et al.
Published: (2026)
ConSiDERS-The-Human Evaluation Framework: Rethinking Human Evaluation for Generative Large Language Models
by: Elangovan, Aparna, et al.
Published: (2024)
by: Elangovan, Aparna, et al.
Published: (2024)
Making Sense Of Distributed Representations With Activation Spectroscopy
by: Reing, Kyle, et al.
Published: (2025)
by: Reing, Kyle, et al.
Published: (2025)
Learning Morphisms with Gauss-Newton Approximation for Growing Networks
by: Lawton, Neal, et al.
Published: (2024)
by: Lawton, Neal, et al.
Published: (2024)
Beyond correlation: The Impact of Human Uncertainty in Measuring the Effectiveness of Automatic Evaluation and LLM-as-a-Judge
by: Elangovan, Aparna, et al.
Published: (2024)
by: Elangovan, Aparna, et al.
Published: (2024)
Exposing Privacy Gaps: Membership Inference Attack on Preference Data for LLM Alignment
by: Feng, Qizhang, et al.
Published: (2024)
by: Feng, Qizhang, et al.
Published: (2024)
Sequential Editing for Lifelong Training of Speech Recognition Models
by: Kulshreshtha, Devang, et al.
Published: (2024)
by: Kulshreshtha, Devang, et al.
Published: (2024)
Estudio y análisis socio-histórico del Control Social
by: Eliú Cardozo Sáez
Published: (2013)
by: Eliú Cardozo Sáez
Published: (2013)
Teaching LLMs to Speak Spectroscopy
by: Ramachandra, Nesar, et al.
Published: (2025)
by: Ramachandra, Nesar, et al.
Published: (2025)
Multi-modal Foundation Model for Cosmological Simulation Data
by: Xia, Bin, et al.
Published: (2025)
by: Xia, Bin, et al.
Published: (2025)
When Retrieval Hurts Code Completion: A Diagnostic Study of Stale Repository Context
by: Weng, Haojun, et al.
Published: (2026)
by: Weng, Haojun, et al.
Published: (2026)
Less Is More? When Dataset Context Hurts LLM-Generated Dataset Descriptions
by: Gan, Lisa-Yao, et al.
Published: (2026)
by: Gan, Lisa-Yao, et al.
Published: (2026)
Wanda++: Pruning Large Language Models via Regional Gradients
by: Yang, Yifan, et al.
Published: (2025)
by: Yang, Yifan, et al.
Published: (2025)
Galaxies and Their Environment at $z \gtrsim 10$ -- I: Primordial Chemical Enrichment, Accretion, Cooling, and Virialization of Gas in Dark Matter Halos
by: Hicks, William M., et al.
Published: (2024)
by: Hicks, William M., et al.
Published: (2024)
Hurt, Baby, Hurt
by: Scott III, William Walter
Published: (2025)
by: Scott III, William Walter
Published: (2025)
A stochastic model for the turbulent ocean heat flux under Arctic sea ice
by: Toppaladoddi, Srikanth, et al.
Published: (2021)
by: Toppaladoddi, Srikanth, et al.
Published: (2021)
Who Gets the Reward, Who Gets the Blame? Evaluation-Aligned Training Signals for Multi-LLM Agents
by: Yang, Chih-Hsuan, et al.
Published: (2025)
by: Yang, Chih-Hsuan, et al.
Published: (2025)
Hijo de tigre... ¿pintito? Perspectivas del tiempo y espacio. Cambios generacionales entre la comunidad cristiana evangélica
by: Angélica Eliú Patiño Reséndiz
Published: (2014)
by: Angélica Eliú Patiño Reséndiz
Published: (2014)
EAIRA: Establishing a Methodology for Evaluating AI Models as Scientific Research Assistants
by: Cappello, Franck, et al.
Published: (2025)
by: Cappello, Franck, et al.
Published: (2025)
SpeechVerse: A Large-scale Generalizable Audio Language Model
by: Das, Nilaksh, et al.
Published: (2024)
by: Das, Nilaksh, et al.
Published: (2024)
The Hurts Don't Hurt Anymore
by: Miller, Betty Davis
Published: (1976)
by: Miller, Betty Davis
Published: (1976)
IdleSpec: Exploiting Idle Time via Speculative Planning for LLM Agents
by: Choi, Daewon, et al.
Published: (2026)
by: Choi, Daewon, et al.
Published: (2026)
When More Retrieval Hurts: Retrieval-Augmented Code Review Generation
by: Meng, Qianru, et al.
Published: (2025)
by: Meng, Qianru, et al.
Published: (2025)
Libraries Alone: Down Under.
by: Sanders, Roy
Published: (1989)
by: Sanders, Roy
Published: (1989)
From Atomistic Models to Machine Learning: Predictive Design of Nanocarbons under Extreme Conditions
by: Yan, Xiaoli, et al.
Published: (2026)
by: Yan, Xiaoli, et al.
Published: (2026)
Similar Items
-
RoPE Distinguishes Neither Positions Nor Tokens in Long Contexts, Provably
by: Du, Yufeng, et al.
Published: (2026) -
Compress, Gather, and Recompute: REFORMing Long-Context Processing in Transformers
by: Song, Woomin, et al.
Published: (2025) -
Adaptive Video Understanding Agent: Enhancing efficiency with dynamic frame sampling and feedback-driven reasoning
by: Jeoung, Sullam, et al.
Published: (2024) -
DCTX-Conformer: Dynamic context carry-over for low latency unified streaming and non-streaming Conformer ASR
by: Huybrechts, Goeric, et al.
Published: (2023) -
SeRA: Self-Reviewing and Alignment of Large Language Models using Implicit Reward Margins
by: Ko, Jongwoo, et al.
Published: (2024)