500xCompressor: Generalized Prompt Compression for Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Zongqian, Su, Yixuan, Collier, Nigel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Prompt Compression for Large Language Models: A Survey
von: Li, Zongqian, et al.
Veröffentlicht: (2024)
von: Li, Zongqian, et al.
Veröffentlicht: (2024)
A Survey on Prompt Tuning
von: Li, Zongqian, et al.
Veröffentlicht: (2025)
von: Li, Zongqian, et al.
Veröffentlicht: (2025)
PT-MoE: An Efficient Finetuning Framework for Integrating Mixture-of-Experts into Prompt Tuning
von: Li, Zongqian, et al.
Veröffentlicht: (2025)
von: Li, Zongqian, et al.
Veröffentlicht: (2025)
Flexi-LoRA with Input-Adaptive Ranks: Efficient Finetuning for Speech and Reasoning Tasks
von: Li, Zongqian, et al.
Veröffentlicht: (2026)
von: Li, Zongqian, et al.
Veröffentlicht: (2026)
ReasonGraph: Visualisation of Reasoning Paths
von: Li, Zongqian, et al.
Veröffentlicht: (2025)
von: Li, Zongqian, et al.
Veröffentlicht: (2025)
Breaking Training Bottlenecks: Effective and Stable Reinforcement Learning for Coding Models
von: Li, Zongqian, et al.
Veröffentlicht: (2026)
von: Li, Zongqian, et al.
Veröffentlicht: (2026)
Unlocking Structure Measuring: Introducing PDD, an Automatic Metric for Positional Discourse Coherence
von: Liu, Yinhong, et al.
Veröffentlicht: (2024)
von: Liu, Yinhong, et al.
Veröffentlicht: (2024)
Large Language Model as Token Compressor and Decompressor
von: Li, Wenbing, et al.
Veröffentlicht: (2026)
von: Li, Wenbing, et al.
Veröffentlicht: (2026)
Attention Instruction: Amplifying Attention in the Middle via Prompting
von: Zhang, Meiru, et al.
Veröffentlicht: (2024)
von: Zhang, Meiru, et al.
Veröffentlicht: (2024)
COFFEE: A Contrastive Oracle-Free Framework for Event Extraction
von: Zhang, Meiru, et al.
Veröffentlicht: (2023)
von: Zhang, Meiru, et al.
Veröffentlicht: (2023)
PiVe: Prompting with Iterative Verification Improving Graph-based Generative Capability of LLMs
von: Han, Jiuzhou, et al.
Veröffentlicht: (2023)
von: Han, Jiuzhou, et al.
Veröffentlicht: (2023)
Scaling Data Difficulty: Improving Coding Models via Reinforcement Learning on Fresh and Challenging Problems
von: Li, Zongqian, et al.
Veröffentlicht: (2026)
von: Li, Zongqian, et al.
Veröffentlicht: (2026)
Conformity in Large Language Models
von: Zhu, Xiaochen, et al.
Veröffentlicht: (2024)
von: Zhu, Xiaochen, et al.
Veröffentlicht: (2024)
iNews: A Multimodal Dataset for Modeling Personalized Affective Responses to News
von: Hu, Tiancheng, et al.
Veröffentlicht: (2025)
von: Hu, Tiancheng, et al.
Veröffentlicht: (2025)
Sparkles: Unlocking Chats Across Multiple Images for Multimodal Instruction-Following Models
von: Huang, Yupan, et al.
Veröffentlicht: (2023)
von: Huang, Yupan, et al.
Veröffentlicht: (2023)
Steer Model beyond Assistant: Controlling System Prompt Strength via Contrastive Decoding
von: Dong, Yijiang River, et al.
Veröffentlicht: (2026)
von: Dong, Yijiang River, et al.
Veröffentlicht: (2026)
Perception Compressor: A Training-Free Prompt Compression Framework in Long Context Scenarios
von: Tang, Jiwei, et al.
Veröffentlicht: (2024)
von: Tang, Jiwei, et al.
Veröffentlicht: (2024)
All Roads Lead to Rome: Graph-Based Confidence Estimation for Large Language Model Reasoning
von: Zhang, Caiqi, et al.
Veröffentlicht: (2025)
von: Zhang, Caiqi, et al.
Veröffentlicht: (2025)
Bi-directional Bias Attribution: Debiasing Large Language Models without Modifying Prompts
von: Lin, Yujie, et al.
Veröffentlicht: (2026)
von: Lin, Yujie, et al.
Veröffentlicht: (2026)
Quantifying the Persona Effect in LLM Simulations
von: Hu, Tiancheng, et al.
Veröffentlicht: (2024)
von: Hu, Tiancheng, et al.
Veröffentlicht: (2024)
Beyond the Final Layer: Intermediate Representations for Better Multilingual Calibration in Large Language Models
von: Zhou, Ej, et al.
Veröffentlicht: (2025)
von: Zhou, Ej, et al.
Veröffentlicht: (2025)
Compressed Sensing for Capability Localization in Large Language Models
von: Bair, Anna, et al.
Veröffentlicht: (2026)
von: Bair, Anna, et al.
Veröffentlicht: (2026)
Generative Language Models Exhibit Social Identity Biases
von: Hu, Tiancheng, et al.
Veröffentlicht: (2023)
von: Hu, Tiancheng, et al.
Veröffentlicht: (2023)
An Empirical Study on Prompt Compression for Large Language Models
von: Zhang, Zheng, et al.
Veröffentlicht: (2025)
von: Zhang, Zheng, et al.
Veröffentlicht: (2025)
The Compressor-Retriever Architecture for Language Model OS
von: Yang, Yuan, et al.
Veröffentlicht: (2024)
von: Yang, Yuan, et al.
Veröffentlicht: (2024)
Retaining Key Information under High Compression Ratios: Query-Guided Compressor for LLMs
von: Cao, Zhiwei, et al.
Veröffentlicht: (2024)
von: Cao, Zhiwei, et al.
Veröffentlicht: (2024)
Aligning with Logic: Measuring, Evaluating and Improving Logical Preference Consistency in Large Language Models
von: Liu, Yinhong, et al.
Veröffentlicht: (2024)
von: Liu, Yinhong, et al.
Veröffentlicht: (2024)
Fairer Preferences Elicit Improved Human-Aligned Large Language Model Judgments
von: Zhou, Han, et al.
Veröffentlicht: (2024)
von: Zhou, Han, et al.
Veröffentlicht: (2024)
SimBench: Benchmarking the Ability of Large Language Models to Simulate Human Behaviors
von: Hu, Tiancheng, et al.
Veröffentlicht: (2025)
von: Hu, Tiancheng, et al.
Veröffentlicht: (2025)
PCToolkit: A Unified Plug-and-Play Prompt Compression Toolkit of Large Language Models
von: Li, Jinyi, et al.
Veröffentlicht: (2024)
von: Li, Jinyi, et al.
Veröffentlicht: (2024)
TransCompressor: LLM-Powered Multimodal Data Compression for Smart Transportation
von: Yang, Huanqi, et al.
Veröffentlicht: (2024)
von: Yang, Huanqi, et al.
Veröffentlicht: (2024)
Time to Revist Exact Match
von: Abbood, Auss, et al.
Veröffentlicht: (2025)
von: Abbood, Auss, et al.
Veröffentlicht: (2025)
SelfCP: Compressing Over-Limit Prompt via the Frozen Large Language Model Itself
von: Gao, Jun, et al.
Veröffentlicht: (2024)
von: Gao, Jun, et al.
Veröffentlicht: (2024)
Dynamic Compressing Prompts for Efficient Inference of Large Language Models
von: Hu, Jinwu, et al.
Veröffentlicht: (2025)
von: Hu, Jinwu, et al.
Veröffentlicht: (2025)
Cmprsr: Abstractive Token-Level Question-Agnostic Prompt Compressor
von: Zakazov, Ivan, et al.
Veröffentlicht: (2025)
von: Zakazov, Ivan, et al.
Veröffentlicht: (2025)
Navigating the Alignment-Calibration Trade-off: A Pareto-Superior Frontier via Model Merging
von: Hu, Tiancheng, et al.
Veröffentlicht: (2025)
von: Hu, Tiancheng, et al.
Veröffentlicht: (2025)
Aligning with Human Judgement: The Role of Pairwise Preference in Large Language Model Evaluators
von: Liu, Yinhong, et al.
Veröffentlicht: (2024)
von: Liu, Yinhong, et al.
Veröffentlicht: (2024)
Prompting Large Language Models for Counterfactual Generation: An Empirical Study
von: Li, Yongqi, et al.
Veröffentlicht: (2023)
von: Li, Yongqi, et al.
Veröffentlicht: (2023)
Can LLM be a Personalized Judge?
von: Dong, Yijiang River, et al.
Veröffentlicht: (2024)
von: Dong, Yijiang River, et al.
Veröffentlicht: (2024)
MaLA-500: Massive Language Adaptation of Large Language Models
von: Lin, Peiqin, et al.
Veröffentlicht: (2024)
von: Lin, Peiqin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Prompt Compression for Large Language Models: A Survey
von: Li, Zongqian, et al.
Veröffentlicht: (2024) -
A Survey on Prompt Tuning
von: Li, Zongqian, et al.
Veröffentlicht: (2025) -
PT-MoE: An Efficient Finetuning Framework for Integrating Mixture-of-Experts into Prompt Tuning
von: Li, Zongqian, et al.
Veröffentlicht: (2025) -
Flexi-LoRA with Input-Adaptive Ranks: Efficient Finetuning for Speech and Reasoning Tasks
von: Li, Zongqian, et al.
Veröffentlicht: (2026) -
ReasonGraph: Visualisation of Reasoning Paths
von: Li, Zongqian, et al.
Veröffentlicht: (2025)