Learning to Compress Prompts with Gist Tokens
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Mu, Jesse, Li, Xiang Lisa, Goodman, Noah |
|---|---|
| Format: | Preprint |
| Publié: |
2023
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Say More with Less: Understanding Prompt Learning Behaviors through Gist Compression
par: Li, Xinze, et autres
Publié: (2024)
par: Li, Xinze, et autres
Publié: (2024)
Compressing Lengthy Context With UltraGist
par: Zhang, Peitian, et autres
Publié: (2024)
par: Zhang, Peitian, et autres
Publié: (2024)
Sentence-Anchored Gist Compression for Long-Context LLMs
par: Tarasov, Dmitrii, et autres
Publié: (2025)
par: Tarasov, Dmitrii, et autres
Publié: (2025)
A Silver Bullet or a Compromise for Full Attention? A Comprehensive Study of Gist Token-based Context Compression
par: Deng, Chenlong, et autres
Publié: (2024)
par: Deng, Chenlong, et autres
Publié: (2024)
GistScore: Learning Better Representations for In-Context Example Selection with Gist Bottlenecks
par: Gupta, Shivanshu, et autres
Publié: (2023)
par: Gupta, Shivanshu, et autres
Publié: (2023)
From Next-Token to Mathematics: The Learning Dynamics of Mathematical Reasoning in Language Models
par: Mishra, Shubhra, et autres
Publié: (2024)
par: Mishra, Shubhra, et autres
Publié: (2024)
UniGist: Towards General and Hardware-aligned Sequence-level Long Context Compression
par: Deng, Chenlong, et autres
Publié: (2025)
par: Deng, Chenlong, et autres
Publié: (2025)
Investigating the Effectiveness of HyperTuning via Gisting
par: Phang, Jason
Publié: (2024)
par: Phang, Jason
Publié: (2024)
Learning to Simulate Human Dialogue
par: Gandhi, Kanishk, et autres
Publié: (2026)
par: Gandhi, Kanishk, et autres
Publié: (2026)
BatchGEMBA: Token-Efficient Machine Translation Evaluation with Batched Prompting and Prompt Compression
par: Larionov, Daniil, et autres
Publié: (2025)
par: Larionov, Daniil, et autres
Publié: (2025)
EmoGist: Efficient In-Context Learning for Visual Emotion Understanding
par: Seoh, Ronald, et autres
Publié: (2025)
par: Seoh, Ronald, et autres
Publié: (2025)
From Token to Token Pair: Efficient Prompt Compression for Large Language Models in Clinical Prediction
par: Zhu, Mingcheng, et autres
Publié: (2026)
par: Zhu, Mingcheng, et autres
Publié: (2026)
An Empirical Study on Prompt Compression for Large Language Models
par: Zhang, Zheng, et autres
Publié: (2025)
par: Zhang, Zheng, et autres
Publié: (2025)
Understand Then Memory: A Cognitive Gist-Driven RAG Framework with Global Semantic Diffusion
par: Zhou, Pengcheng, et autres
Publié: (2026)
par: Zhou, Pengcheng, et autres
Publié: (2026)
Discrete Prompt Compression with Reinforcement Learning
par: Jung, Hoyoun, et autres
Publié: (2023)
par: Jung, Hoyoun, et autres
Publié: (2023)
SciGisPy: a Novel Metric for Biomedical Text Simplification via Gist Inference Score
par: Lyu, Chen, et autres
Publié: (2024)
par: Lyu, Chen, et autres
Publié: (2024)
Large Language Model Reasoning Failures
par: Song, Peiyang, et autres
Publié: (2026)
par: Song, Peiyang, et autres
Publié: (2026)
PromptCL: Improving Event Representation via Prompt Template and Contrastive Learning
par: Feng, Yubo, et autres
Publié: (2024)
par: Feng, Yubo, et autres
Publié: (2024)
Automated Statistical Model Discovery with Language Models
par: Li, Michael Y., et autres
Publié: (2024)
par: Li, Michael Y., et autres
Publié: (2024)
More Tokens, Lower Precision: Towards the Optimal Token-Precision Trade-off in KV Cache Compression
par: Zhang, Jiebin, et autres
Publié: (2024)
par: Zhang, Jiebin, et autres
Publié: (2024)
Behavior-Equivalent Token: Single-Token Replacement for Long Prompts in LLMs
par: Dong, Jiancheng, et autres
Publié: (2025)
par: Dong, Jiancheng, et autres
Publié: (2025)
Words that make SENSE: Sensorimotor Norms in Learned Lexical Token Representations
par: Gupta, Abhinav, et autres
Publié: (2026)
par: Gupta, Abhinav, et autres
Publié: (2026)
Evaluating and Optimizing Educational Content with Large Language Model Judgments
par: He-Yueya, Joy, et autres
Publié: (2024)
par: He-Yueya, Joy, et autres
Publié: (2024)
A Human-Inspired Reading Agent with Gist Memory of Very Long Contexts
par: Lee, Kuang-Huei, et autres
Publié: (2024)
par: Lee, Kuang-Huei, et autres
Publié: (2024)
Prompt Compression for Large Language Models: A Survey
par: Li, Zongqian, et autres
Publié: (2024)
par: Li, Zongqian, et autres
Publié: (2024)
Self-Supervised Alignment with Mutual Information: Learning to Follow Principles without Preference Labels
par: Fränken, Jan-Philipp, et autres
Publié: (2024)
par: Fränken, Jan-Philipp, et autres
Publié: (2024)
TokenSkip: Controllable Chain-of-Thought Compression in LLMs
par: Xia, Heming, et autres
Publié: (2025)
par: Xia, Heming, et autres
Publié: (2025)
Incorporating Token Usage into Prompting Strategy Evaluation
par: Sypherd, Chris, et autres
Publié: (2025)
par: Sypherd, Chris, et autres
Publié: (2025)
DAST: Context-Aware Compression in LLMs via Dynamic Allocation of Soft Tokens
par: Chen, Shaoshen, et autres
Publié: (2025)
par: Chen, Shaoshen, et autres
Publié: (2025)
Beyond Text Compression: Evaluating Tokenizers Across Scales
par: Lotz, Jonas F., et autres
Publié: (2025)
par: Lotz, Jonas F., et autres
Publié: (2025)
Interpretability at Scale: Identifying Causal Mechanisms in Alpaca
par: Wu, Zhengxuan, et autres
Publié: (2023)
par: Wu, Zhengxuan, et autres
Publié: (2023)
Endless Terminals: Scaling RL Environments for Terminal Agents
par: Gandhi, Kanishk, et autres
Publié: (2026)
par: Gandhi, Kanishk, et autres
Publié: (2026)
Breaking Token Into Concepts: Exploring Extreme Compression in Token Representation Via Compositional Shared Semantics
par: R V, Kavin, et autres
Publié: (2025)
par: R V, Kavin, et autres
Publié: (2025)
Multi-word Tokenization for Sequence Compression
par: Gee, Leonidas, et autres
Publié: (2024)
par: Gee, Leonidas, et autres
Publié: (2024)
Broken Tokens? Your Language Model can Secretly Handle Non-Canonical Tokenizations
par: Zheng, Brian Siyuan, et autres
Publié: (2025)
par: Zheng, Brian Siyuan, et autres
Publié: (2025)
Learning to Compress Prompt in Natural Language Formats
par: Chuang, Yu-Neng, et autres
Publié: (2024)
par: Chuang, Yu-Neng, et autres
Publié: (2024)
500xCompressor: Generalized Prompt Compression for Large Language Models
par: Li, Zongqian, et autres
Publié: (2024)
par: Li, Zongqian, et autres
Publié: (2024)
Lossless Token Sequence Compression via Meta-Tokens
par: Harvill, John, et autres
Publié: (2025)
par: Harvill, John, et autres
Publié: (2025)
Detecting Overflow in Compressed Token Representations for Retrieval-Augmented Generation
par: Belikova, Julia, et autres
Publié: (2026)
par: Belikova, Julia, et autres
Publié: (2026)
APLe: Token-Wise Adaptive for Multi-Modal Prompt Learning
par: Cao, Guiming, et autres
Publié: (2024)
par: Cao, Guiming, et autres
Publié: (2024)
Documents similaires
-
Say More with Less: Understanding Prompt Learning Behaviors through Gist Compression
par: Li, Xinze, et autres
Publié: (2024) -
Compressing Lengthy Context With UltraGist
par: Zhang, Peitian, et autres
Publié: (2024) -
Sentence-Anchored Gist Compression for Long-Context LLMs
par: Tarasov, Dmitrii, et autres
Publié: (2025) -
A Silver Bullet or a Compromise for Full Attention? A Comprehensive Study of Gist Token-based Context Compression
par: Deng, Chenlong, et autres
Publié: (2024) -
GistScore: Learning Better Representations for In-Context Example Selection with Gist Bottlenecks
par: Gupta, Shivanshu, et autres
Publié: (2023)