ProCut: LLM Prompt Compression via Attribution Estimation
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Zhentao, Li, Fengyi, Chen, Albert, Wang, Xiaofeng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SPA-RL: Reinforcing LLM Agents via Stepwise Progress Attribution
by: Wang, Hanlin, et al.
Published: (2025)
by: Wang, Hanlin, et al.
Published: (2025)
JoPA:Explaining Large Language Model's Generation via Joint Prompt Attribution
by: Chang, Yurui, et al.
Published: (2024)
by: Chang, Yurui, et al.
Published: (2024)
Prompt Compression with Context-Aware Sentence Encoding for Fast and Improved LLM Inference
by: Liskavets, Barys, et al.
Published: (2024)
by: Liskavets, Barys, et al.
Published: (2024)
BP-Seg: A graphical model approach to unsupervised and non-contiguous text segmentation using belief propagation
by: Li, Fengyi, et al.
Published: (2025)
by: Li, Fengyi, et al.
Published: (2025)
LLM Prompt Duel Optimizer: Efficient Label-Free Prompt Optimization
by: Wu, Yuanchen, et al.
Published: (2025)
by: Wu, Yuanchen, et al.
Published: (2025)
LongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression
by: Jiang, Huiqiang, et al.
Published: (2023)
by: Jiang, Huiqiang, et al.
Published: (2023)
ProPD: Dynamic Token Tree Pruning and Generation for LLM Parallel Decoding
by: Zhong, Shuzhang, et al.
Published: (2024)
by: Zhong, Shuzhang, et al.
Published: (2024)
PocketLLM: Ultimate Compression of Large Language Models via Meta Networks
by: Tian, Ye, et al.
Published: (2025)
by: Tian, Ye, et al.
Published: (2025)
Entropy Law: The Story Behind Data Compression and LLM Performance
by: Yin, Mingjia, et al.
Published: (2024)
by: Yin, Mingjia, et al.
Published: (2024)
LLM as GNN: Graph Vocabulary Learning for Text-Attributed Graph Foundation Models
by: Zhu, Xi, et al.
Published: (2025)
by: Zhu, Xi, et al.
Published: (2025)
Attribution analysis of legal language as used by LLM
by: Belew, Richard K.
Published: (2025)
by: Belew, Richard K.
Published: (2025)
Prompt-SAW: Leveraging Relation-Aware Graphs for Textual Prompt Compression
by: Ali, Muhammad Asif, et al.
Published: (2024)
by: Ali, Muhammad Asif, et al.
Published: (2024)
Better Prompt Compression Without Multi-Layer Perceptrons
by: Honig, Edouardo, et al.
Published: (2025)
by: Honig, Edouardo, et al.
Published: (2025)
Characterizing Prompt Compression Methods for Long Context Inference
by: Jha, Siddharth, et al.
Published: (2024)
by: Jha, Siddharth, et al.
Published: (2024)
SIG: Speaker Identification in Literature via Prompt-Based Generation
by: Su, Zhenlin, et al.
Published: (2023)
by: Su, Zhenlin, et al.
Published: (2023)
PromptFix: Few-shot Backdoor Removal via Adversarial Prompt Tuning
by: Zhang, Tianrong, et al.
Published: (2024)
by: Zhang, Tianrong, et al.
Published: (2024)
Learning to Compress Prompt in Natural Language Formats
by: Chuang, Yu-Neng, et al.
Published: (2024)
by: Chuang, Yu-Neng, et al.
Published: (2024)
QET: Enhancing Quantized LLM Parameters and KV cache Compression through Element Substitution and Residual Clustering
by: Wang, Yanshu, et al.
Published: (2024)
by: Wang, Yanshu, et al.
Published: (2024)
ProRefine: Inference-Time Prompt Refinement with Textual Feedback
by: Pandita, Deepak, et al.
Published: (2025)
by: Pandita, Deepak, et al.
Published: (2025)
Where Did It Go Wrong? Attributing Undesirable LLM Behaviors via Representation Gradient Tracing
by: Li, Zhe, et al.
Published: (2025)
by: Li, Zhe, et al.
Published: (2025)
Prompt Candidates, then Distill: A Teacher-Student Framework for LLM-driven Data Annotation
by: Xia, Mingxuan, et al.
Published: (2025)
by: Xia, Mingxuan, et al.
Published: (2025)
PolarQuant: Optimal Gaussian Weight Quantization via Hadamard Rotation for LLM Compression
by: Vicentino, Caio
Published: (2026)
by: Vicentino, Caio
Published: (2026)
Extracting Prompts by Inverting LLM Outputs
by: Zhang, Collin, et al.
Published: (2024)
by: Zhang, Collin, et al.
Published: (2024)
LLM In-Context Recall is Prompt Dependent
by: Machlab, Daniel, et al.
Published: (2024)
by: Machlab, Daniel, et al.
Published: (2024)
Prompt Curriculum Learning for Efficient LLM Post-Training
by: Gao, Zhaolin, et al.
Published: (2025)
by: Gao, Zhaolin, et al.
Published: (2025)
TACO-RL: Task Aware Prompt Compression Optimization with Reinforcement Learning
by: Shandilya, Shivam, et al.
Published: (2024)
by: Shandilya, Shivam, et al.
Published: (2024)
Prompt Optimization via Adversarial In-Context Learning
by: Do, Xuan Long, et al.
Published: (2023)
by: Do, Xuan Long, et al.
Published: (2023)
Hardware-Aware Parallel Prompt Decoding for Memory-Efficient Acceleration of LLM Inference
by: Chen, Hao Mark, et al.
Published: (2024)
by: Chen, Hao Mark, et al.
Published: (2024)
Compress the Context, Keep the Commitments: A Formal Framework for Verifiable LLM Context Compression
by: Trukhina, Natalia, et al.
Published: (2026)
by: Trukhina, Natalia, et al.
Published: (2026)
Does Prompt Formatting Have Any Impact on LLM Performance?
by: He, Jia, et al.
Published: (2024)
by: He, Jia, et al.
Published: (2024)
DynSplit-KV: Dynamic Semantic Splitting for KVCache Compression in Efficient Long-Context LLM Inference
by: Ye, Jiancai, et al.
Published: (2026)
by: Ye, Jiancai, et al.
Published: (2026)
LLMLingua-2: Data Distillation for Efficient and Faithful Task-Agnostic Prompt Compression
by: Pan, Zhuoshi, et al.
Published: (2024)
by: Pan, Zhuoshi, et al.
Published: (2024)
Adapting LLMs for Efficient Context Processing through Soft Prompt Compression
by: Wang, Cangqing, et al.
Published: (2024)
by: Wang, Cangqing, et al.
Published: (2024)
Dialectical Behavior Therapy Approach to LLM Prompting
by: Vitman, Oxana, et al.
Published: (2024)
by: Vitman, Oxana, et al.
Published: (2024)
Rethinking LLM Memorization through the Lens of Adversarial Compression
by: Schwarzschild, Avi, et al.
Published: (2024)
by: Schwarzschild, Avi, et al.
Published: (2024)
Prompt Exploration with Prompt Regression
by: Feffer, Michael, et al.
Published: (2024)
by: Feffer, Michael, et al.
Published: (2024)
RocketKV: Accelerating Long-Context LLM Inference via Two-Stage KV Cache Compression
by: Behnam, Payman, et al.
Published: (2025)
by: Behnam, Payman, et al.
Published: (2025)
Question-Aware Knowledge Graph Prompting for Enhancing Large Language Models
by: Liu, Haochen, et al.
Published: (2025)
by: Liu, Haochen, et al.
Published: (2025)
Mechanistic Data Attribution: Tracing the Training Origins of Interpretable LLM Units
by: Chen, Jianhui, et al.
Published: (2026)
by: Chen, Jianhui, et al.
Published: (2026)
MatryoshkaKV: Adaptive KV Compression via Trainable Orthogonal Projection
by: Lin, Bokai, et al.
Published: (2024)
by: Lin, Bokai, et al.
Published: (2024)
Similar Items
-
SPA-RL: Reinforcing LLM Agents via Stepwise Progress Attribution
by: Wang, Hanlin, et al.
Published: (2025) -
JoPA:Explaining Large Language Model's Generation via Joint Prompt Attribution
by: Chang, Yurui, et al.
Published: (2024) -
Prompt Compression with Context-Aware Sentence Encoding for Fast and Improved LLM Inference
by: Liskavets, Barys, et al.
Published: (2024) -
BP-Seg: A graphical model approach to unsupervised and non-contiguous text segmentation using belief propagation
by: Li, Fengyi, et al.
Published: (2025) -
LLM Prompt Duel Optimizer: Efficient Label-Free Prompt Optimization
by: Wu, Yuanchen, et al.
Published: (2025)