Gespeichert in:
| Hauptverfasser: | Gao, Yaxin, Lu, Yao, Zhang, Zongfei, Nie, Jiaqi, Yu, Shanqing, Xuan, Qi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2509.13723 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
From LLM-anation to LLM-orchestrator: Coordinating Small Models for Data Labeling
von: Lu, Yao, et al.
Veröffentlicht: (2025)
von: Lu, Yao, et al.
Veröffentlicht: (2025)
CCF: A Context Compression Framework for Efficient Long-Sequence Language Modeling
von: Li, Wenhao, et al.
Veröffentlicht: (2025)
von: Li, Wenhao, et al.
Veröffentlicht: (2025)
Decoupled Reasoning with Implicit Fact Tokens (DRIFT): A Dual-Model Framework for Efficient Long-Context Inference
von: Xie, Wenxuan, et al.
Veröffentlicht: (2026)
von: Xie, Wenxuan, et al.
Veröffentlicht: (2026)
VTC-R1: Vision-Text Compression for Efficient Long-Context Reasoning
von: Wang, Yibo, et al.
Veröffentlicht: (2026)
von: Wang, Yibo, et al.
Veröffentlicht: (2026)
LongAttnComp: Cross-Family Context Compression for Long-Context Reasoning
von: Ji, Mengmeng, et al.
Veröffentlicht: (2026)
von: Ji, Mengmeng, et al.
Veröffentlicht: (2026)
LoRALib: A Standardized Benchmark for Evaluating LoRA-MoE Methods
von: Wang, Shaoheng, et al.
Veröffentlicht: (2025)
von: Wang, Shaoheng, et al.
Veröffentlicht: (2025)
Perception Compressor: A Training-Free Prompt Compression Framework in Long Context Scenarios
von: Tang, Jiwei, et al.
Veröffentlicht: (2024)
von: Tang, Jiwei, et al.
Veröffentlicht: (2024)
Legal Mathematical Reasoning with LLMs: Procedural Alignment through Two-Stage Reinforcement Learning
von: Zhang, Kepu, et al.
Veröffentlicht: (2025)
von: Zhang, Kepu, et al.
Veröffentlicht: (2025)
Dual Knowledge-Enhanced Two-Stage Reasoner for Multimodal Dialog Systems
von: Chen, Xiaolin, et al.
Veröffentlicht: (2025)
von: Chen, Xiaolin, et al.
Veröffentlicht: (2025)
HalluClean: A Unified Framework to Combat Hallucinations in LLMs
von: Zhao, Yaxin, et al.
Veröffentlicht: (2025)
von: Zhao, Yaxin, et al.
Veröffentlicht: (2025)
QUITO: Accelerating Long-Context Reasoning through Query-Guided Context Compression
von: Wang, Wenshan, et al.
Veröffentlicht: (2024)
von: Wang, Wenshan, et al.
Veröffentlicht: (2024)
FastCuRL: Curriculum Reinforcement Learning with Stage-wise Context Scaling for Efficient Training R1-like Reasoning Models
von: Song, Mingyang, et al.
Veröffentlicht: (2025)
von: Song, Mingyang, et al.
Veröffentlicht: (2025)
RocketKV: Accelerating Long-Context LLM Inference via Two-Stage KV Cache Compression
von: Behnam, Payman, et al.
Veröffentlicht: (2025)
von: Behnam, Payman, et al.
Veröffentlicht: (2025)
Text-to-SQL as Dual-State Reasoning: Integrating Adaptive Context and Progressive Generation
von: Hao, Zhifeng, et al.
Veröffentlicht: (2025)
von: Hao, Zhifeng, et al.
Veröffentlicht: (2025)
ICXML: An In-Context Learning Framework for Zero-Shot Extreme Multi-Label Classification
von: Zhu, Yaxin, et al.
Veröffentlicht: (2023)
von: Zhu, Yaxin, et al.
Veröffentlicht: (2023)
LongFlow: Efficient KV Cache Compression for Reasoning Models
von: Su, Yi, et al.
Veröffentlicht: (2026)
von: Su, Yi, et al.
Veröffentlicht: (2026)
LooGLE: Can Long-Context Language Models Understand Long Contexts?
von: Li, Jiaqi, et al.
Veröffentlicht: (2023)
von: Li, Jiaqi, et al.
Veröffentlicht: (2023)
Efficient Prompt Compression with Evaluator Heads for Long-Context Transformer Inference
von: Fei, Weizhi, et al.
Veröffentlicht: (2025)
von: Fei, Weizhi, et al.
Veröffentlicht: (2025)
Multipole Attention for Efficient Long Context Reasoning
von: Hooper, Coleman, et al.
Veröffentlicht: (2025)
von: Hooper, Coleman, et al.
Veröffentlicht: (2025)
DELTA: Dynamic Layer-Aware Token Attention for Efficient Long-Context Reasoning
von: Zarch, Hossein Entezari, et al.
Veröffentlicht: (2025)
von: Zarch, Hossein Entezari, et al.
Veröffentlicht: (2025)
Every Attention Matters: An Efficient Hybrid Architecture for Long-Context Reasoning
von: Ling Team, et al.
Veröffentlicht: (2025)
von: Ling Team, et al.
Veröffentlicht: (2025)
Long Context Compression with Activation Beacon
von: Zhang, Peitian, et al.
Veröffentlicht: (2024)
von: Zhang, Peitian, et al.
Veröffentlicht: (2024)
Dynamic Long Context Reasoning over Compressed Memory via End-to-End Reinforcement Learning
von: Chen, Zhuoen, et al.
Veröffentlicht: (2026)
von: Chen, Zhuoen, et al.
Veröffentlicht: (2026)
BRIEF-Pro: Universal Context Compression with Short-to-Long Synthesis for Fast and Accurate Multi-Hop Reasoning
von: Gu, Jia-Chen, et al.
Veröffentlicht: (2025)
von: Gu, Jia-Chen, et al.
Veröffentlicht: (2025)
UniICL: An Efficient Unified Framework Unifying Compression, Selection, and Generation
von: Gao, Jun, et al.
Veröffentlicht: (2024)
von: Gao, Jun, et al.
Veröffentlicht: (2024)
TriAttention: Efficient Long Reasoning with Trigonometric KV Compression
von: Mao, Weian, et al.
Veröffentlicht: (2026)
von: Mao, Weian, et al.
Veröffentlicht: (2026)
HyLRA: Hybrid Layer Reuse Attention for Efficient Long-Context Inference
von: Ai, Xuan, et al.
Veröffentlicht: (2026)
von: Ai, Xuan, et al.
Veröffentlicht: (2026)
Compressing KV Cache for Long-Context LLM Inference with Inter-Layer Attention Similarity
von: Ma, Da, et al.
Veröffentlicht: (2024)
von: Ma, Da, et al.
Veröffentlicht: (2024)
LongReason: A Synthetic Long-Context Reasoning Benchmark via Context Expansion
von: Ling, Zhan, et al.
Veröffentlicht: (2025)
von: Ling, Zhan, et al.
Veröffentlicht: (2025)
LIFT: A Novel Framework for Enhancing Long-Context Understanding of LLMs via Long Input Fine-Tuning
von: Mao, Yansheng, et al.
Veröffentlicht: (2025)
von: Mao, Yansheng, et al.
Veröffentlicht: (2025)
Context Memorization for Efficient Long Context Generation
von: Okoshi, Yasuyuki, et al.
Veröffentlicht: (2026)
von: Okoshi, Yasuyuki, et al.
Veröffentlicht: (2026)
ATTNPO: Attention-Guided Process Supervision for Efficient Reasoning
von: Nie, Shuaiyi, et al.
Veröffentlicht: (2026)
von: Nie, Shuaiyi, et al.
Veröffentlicht: (2026)
A Self-Evolving Framework for Efficient Terminal Agents via Observational Context Compression
von: Ren, Jincheng, et al.
Veröffentlicht: (2026)
von: Ren, Jincheng, et al.
Veröffentlicht: (2026)
Reasoning Path Compression: Compressing Generation Trajectories for Efficient LLM Reasoning
von: Song, Jiwon, et al.
Veröffentlicht: (2025)
von: Song, Jiwon, et al.
Veröffentlicht: (2025)
LongCodeZip: Compress Long Context for Code Language Models
von: Shi, Yuling, et al.
Veröffentlicht: (2025)
von: Shi, Yuling, et al.
Veröffentlicht: (2025)
UniGist: Towards General and Hardware-aligned Sequence-level Long Context Compression
von: Deng, Chenlong, et al.
Veröffentlicht: (2025)
von: Deng, Chenlong, et al.
Veröffentlicht: (2025)
From Harm to Help: Turning Reasoning In-Context Demos into Assets for Reasoning LMs
von: Wang, Haonan, et al.
Veröffentlicht: (2025)
von: Wang, Haonan, et al.
Veröffentlicht: (2025)
ATACompressor: Adaptive Task-Aware Compression for Efficient Long-Context Processing in LLMs
von: Li, Xuancheng, et al.
Veröffentlicht: (2026)
von: Li, Xuancheng, et al.
Veröffentlicht: (2026)
NestedKV: Nested Memory Routing for Long-Context KV Cache Compression
von: Chen, Hong, et al.
Veröffentlicht: (2026)
von: Chen, Hong, et al.
Veröffentlicht: (2026)
ReCoG: Relational and Compact Context Graph Learning for Few-shot Molecular Property Prediction
von: Wang, Zeyu, et al.
Veröffentlicht: (2026)
von: Wang, Zeyu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
From LLM-anation to LLM-orchestrator: Coordinating Small Models for Data Labeling
von: Lu, Yao, et al.
Veröffentlicht: (2025) -
CCF: A Context Compression Framework for Efficient Long-Sequence Language Modeling
von: Li, Wenhao, et al.
Veröffentlicht: (2025) -
Decoupled Reasoning with Implicit Fact Tokens (DRIFT): A Dual-Model Framework for Efficient Long-Context Inference
von: Xie, Wenxuan, et al.
Veröffentlicht: (2026) -
VTC-R1: Vision-Text Compression for Efficient Long-Context Reasoning
von: Wang, Yibo, et al.
Veröffentlicht: (2026) -
LongAttnComp: Cross-Family Context Compression for Long-Context Reasoning
von: Ji, Mengmeng, et al.
Veröffentlicht: (2026)