Learning to Compress Prompts with Gist Tokens
Fuente:
arXiv
Saved in:
| Main Authors: | Mu, Jesse, Li, Xiang Lisa, Goodman, Noah |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Say More with Less: Understanding Prompt Learning Behaviors through Gist Compression
by: Li, Xinze, et al.
Published: (2024)
by: Li, Xinze, et al.
Published: (2024)
Compressing Lengthy Context With UltraGist
by: Zhang, Peitian, et al.
Published: (2024)
by: Zhang, Peitian, et al.
Published: (2024)
Sentence-Anchored Gist Compression for Long-Context LLMs
by: Tarasov, Dmitrii, et al.
Published: (2025)
by: Tarasov, Dmitrii, et al.
Published: (2025)
A Silver Bullet or a Compromise for Full Attention? A Comprehensive Study of Gist Token-based Context Compression
by: Deng, Chenlong, et al.
Published: (2024)
by: Deng, Chenlong, et al.
Published: (2024)
GistScore: Learning Better Representations for In-Context Example Selection with Gist Bottlenecks
by: Gupta, Shivanshu, et al.
Published: (2023)
by: Gupta, Shivanshu, et al.
Published: (2023)
From Next-Token to Mathematics: The Learning Dynamics of Mathematical Reasoning in Language Models
by: Mishra, Shubhra, et al.
Published: (2024)
by: Mishra, Shubhra, et al.
Published: (2024)
UniGist: Towards General and Hardware-aligned Sequence-level Long Context Compression
by: Deng, Chenlong, et al.
Published: (2025)
by: Deng, Chenlong, et al.
Published: (2025)
Investigating the Effectiveness of HyperTuning via Gisting
by: Phang, Jason
Published: (2024)
by: Phang, Jason
Published: (2024)
Learning to Simulate Human Dialogue
by: Gandhi, Kanishk, et al.
Published: (2026)
by: Gandhi, Kanishk, et al.
Published: (2026)
BatchGEMBA: Token-Efficient Machine Translation Evaluation with Batched Prompting and Prompt Compression
by: Larionov, Daniil, et al.
Published: (2025)
by: Larionov, Daniil, et al.
Published: (2025)
EmoGist: Efficient In-Context Learning for Visual Emotion Understanding
by: Seoh, Ronald, et al.
Published: (2025)
by: Seoh, Ronald, et al.
Published: (2025)
From Token to Token Pair: Efficient Prompt Compression for Large Language Models in Clinical Prediction
by: Zhu, Mingcheng, et al.
Published: (2026)
by: Zhu, Mingcheng, et al.
Published: (2026)
An Empirical Study on Prompt Compression for Large Language Models
by: Zhang, Zheng, et al.
Published: (2025)
by: Zhang, Zheng, et al.
Published: (2025)
Understand Then Memory: A Cognitive Gist-Driven RAG Framework with Global Semantic Diffusion
by: Zhou, Pengcheng, et al.
Published: (2026)
by: Zhou, Pengcheng, et al.
Published: (2026)
Discrete Prompt Compression with Reinforcement Learning
by: Jung, Hoyoun, et al.
Published: (2023)
by: Jung, Hoyoun, et al.
Published: (2023)
SciGisPy: a Novel Metric for Biomedical Text Simplification via Gist Inference Score
by: Lyu, Chen, et al.
Published: (2024)
by: Lyu, Chen, et al.
Published: (2024)
Large Language Model Reasoning Failures
by: Song, Peiyang, et al.
Published: (2026)
by: Song, Peiyang, et al.
Published: (2026)
PromptCL: Improving Event Representation via Prompt Template and Contrastive Learning
by: Feng, Yubo, et al.
Published: (2024)
by: Feng, Yubo, et al.
Published: (2024)
Automated Statistical Model Discovery with Language Models
by: Li, Michael Y., et al.
Published: (2024)
by: Li, Michael Y., et al.
Published: (2024)
More Tokens, Lower Precision: Towards the Optimal Token-Precision Trade-off in KV Cache Compression
by: Zhang, Jiebin, et al.
Published: (2024)
by: Zhang, Jiebin, et al.
Published: (2024)
Behavior-Equivalent Token: Single-Token Replacement for Long Prompts in LLMs
by: Dong, Jiancheng, et al.
Published: (2025)
by: Dong, Jiancheng, et al.
Published: (2025)
Words that make SENSE: Sensorimotor Norms in Learned Lexical Token Representations
by: Gupta, Abhinav, et al.
Published: (2026)
by: Gupta, Abhinav, et al.
Published: (2026)
Evaluating and Optimizing Educational Content with Large Language Model Judgments
by: He-Yueya, Joy, et al.
Published: (2024)
by: He-Yueya, Joy, et al.
Published: (2024)
A Human-Inspired Reading Agent with Gist Memory of Very Long Contexts
by: Lee, Kuang-Huei, et al.
Published: (2024)
by: Lee, Kuang-Huei, et al.
Published: (2024)
Prompt Compression for Large Language Models: A Survey
by: Li, Zongqian, et al.
Published: (2024)
by: Li, Zongqian, et al.
Published: (2024)
Self-Supervised Alignment with Mutual Information: Learning to Follow Principles without Preference Labels
by: Fränken, Jan-Philipp, et al.
Published: (2024)
by: Fränken, Jan-Philipp, et al.
Published: (2024)
TokenSkip: Controllable Chain-of-Thought Compression in LLMs
by: Xia, Heming, et al.
Published: (2025)
by: Xia, Heming, et al.
Published: (2025)
Incorporating Token Usage into Prompting Strategy Evaluation
by: Sypherd, Chris, et al.
Published: (2025)
by: Sypherd, Chris, et al.
Published: (2025)
DAST: Context-Aware Compression in LLMs via Dynamic Allocation of Soft Tokens
by: Chen, Shaoshen, et al.
Published: (2025)
by: Chen, Shaoshen, et al.
Published: (2025)
Beyond Text Compression: Evaluating Tokenizers Across Scales
by: Lotz, Jonas F., et al.
Published: (2025)
by: Lotz, Jonas F., et al.
Published: (2025)
Interpretability at Scale: Identifying Causal Mechanisms in Alpaca
by: Wu, Zhengxuan, et al.
Published: (2023)
by: Wu, Zhengxuan, et al.
Published: (2023)
Endless Terminals: Scaling RL Environments for Terminal Agents
by: Gandhi, Kanishk, et al.
Published: (2026)
by: Gandhi, Kanishk, et al.
Published: (2026)
Breaking Token Into Concepts: Exploring Extreme Compression in Token Representation Via Compositional Shared Semantics
by: R V, Kavin, et al.
Published: (2025)
by: R V, Kavin, et al.
Published: (2025)
Multi-word Tokenization for Sequence Compression
by: Gee, Leonidas, et al.
Published: (2024)
by: Gee, Leonidas, et al.
Published: (2024)
Broken Tokens? Your Language Model can Secretly Handle Non-Canonical Tokenizations
by: Zheng, Brian Siyuan, et al.
Published: (2025)
by: Zheng, Brian Siyuan, et al.
Published: (2025)
Learning to Compress Prompt in Natural Language Formats
by: Chuang, Yu-Neng, et al.
Published: (2024)
by: Chuang, Yu-Neng, et al.
Published: (2024)
500xCompressor: Generalized Prompt Compression for Large Language Models
by: Li, Zongqian, et al.
Published: (2024)
by: Li, Zongqian, et al.
Published: (2024)
Lossless Token Sequence Compression via Meta-Tokens
by: Harvill, John, et al.
Published: (2025)
by: Harvill, John, et al.
Published: (2025)
Detecting Overflow in Compressed Token Representations for Retrieval-Augmented Generation
by: Belikova, Julia, et al.
Published: (2026)
by: Belikova, Julia, et al.
Published: (2026)
APLe: Token-Wise Adaptive for Multi-Modal Prompt Learning
by: Cao, Guiming, et al.
Published: (2024)
by: Cao, Guiming, et al.
Published: (2024)
Similar Items
-
Say More with Less: Understanding Prompt Learning Behaviors through Gist Compression
by: Li, Xinze, et al.
Published: (2024) -
Compressing Lengthy Context With UltraGist
by: Zhang, Peitian, et al.
Published: (2024) -
Sentence-Anchored Gist Compression for Long-Context LLMs
by: Tarasov, Dmitrii, et al.
Published: (2025) -
A Silver Bullet or a Compromise for Full Attention? A Comprehensive Study of Gist Token-based Context Compression
by: Deng, Chenlong, et al.
Published: (2024) -
GistScore: Learning Better Representations for In-Context Example Selection with Gist Bottlenecks
by: Gupta, Shivanshu, et al.
Published: (2023)