Compressing Lengthy Context With UltraGist
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Peitian, Liu, Zheng, Xiao, Shitao, Shao, Ninglu, Ye, Qiwei, Dou, Zhicheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Long Context Compression with Activation Beacon
von: Zhang, Peitian, et al.
Veröffentlicht: (2024)
von: Zhang, Peitian, et al.
Veröffentlicht: (2024)
Extending Llama-3's Context Ten-Fold Overnight
von: Zhang, Peitian, et al.
Veröffentlicht: (2024)
von: Zhang, Peitian, et al.
Veröffentlicht: (2024)
Extensible Embedding: A Flexible Multipler For LLM's Context Length
von: Shao, Ninglu, et al.
Veröffentlicht: (2024)
von: Shao, Ninglu, et al.
Veröffentlicht: (2024)
Flexibly Scaling Large Language Models Contexts Through Extensible Tokenization
von: Shao, Ninglu, et al.
Veröffentlicht: (2024)
von: Shao, Ninglu, et al.
Veröffentlicht: (2024)
Lighter And Better: Towards Flexible Context Adaptation For Retrieval Augmented Generation
von: Liu, Zheng, et al.
Veröffentlicht: (2024)
von: Liu, Zheng, et al.
Veröffentlicht: (2024)
UniGist: Towards General and Hardware-aligned Sequence-level Long Context Compression
von: Deng, Chenlong, et al.
Veröffentlicht: (2025)
von: Deng, Chenlong, et al.
Veröffentlicht: (2025)
Boosting Long-Context Management via Query-Guided Activation Refilling
von: Qian, Hongjin, et al.
Veröffentlicht: (2024)
von: Qian, Hongjin, et al.
Veröffentlicht: (2024)
A Silver Bullet or a Compromise for Full Attention? A Comprehensive Study of Gist Token-based Context Compression
von: Deng, Chenlong, et al.
Veröffentlicht: (2024)
von: Deng, Chenlong, et al.
Veröffentlicht: (2024)
AdmTree: Compressing Lengthy Context with Adaptive Semantic Trees
von: Li, Yangning, et al.
Veröffentlicht: (2025)
von: Li, Yangning, et al.
Veröffentlicht: (2025)
Are Long-LLMs A Necessity For Long-Context Tasks?
von: Qian, Hongjin, et al.
Veröffentlicht: (2024)
von: Qian, Hongjin, et al.
Veröffentlicht: (2024)
Sentence-Anchored Gist Compression for Long-Context LLMs
von: Tarasov, Dmitrii, et al.
Veröffentlicht: (2025)
von: Tarasov, Dmitrii, et al.
Veröffentlicht: (2025)
MemoRAG: Boosting Long Context Processing with Global Memory-Enhanced Retrieval Augmentation
von: Qian, Hongjin, et al.
Veröffentlicht: (2024)
von: Qian, Hongjin, et al.
Veröffentlicht: (2024)
M3-Embedding: Multi-Linguality, Multi-Functionality, Multi-Granularity Text Embeddings Through Self-Knowledge Distillation
von: Chen, Jianlv, et al.
Veröffentlicht: (2024)
von: Chen, Jianlv, et al.
Veröffentlicht: (2024)
Learning to Compress Prompts with Gist Tokens
von: Mu, Jesse, et al.
Veröffentlicht: (2023)
von: Mu, Jesse, et al.
Veröffentlicht: (2023)
GistScore: Learning Better Representations for In-Context Example Selection with Gist Bottlenecks
von: Gupta, Shivanshu, et al.
Veröffentlicht: (2023)
von: Gupta, Shivanshu, et al.
Veröffentlicht: (2023)
C-Pack: Packed Resources For General Chinese Embeddings
von: Xiao, Shitao, et al.
Veröffentlicht: (2023)
von: Xiao, Shitao, et al.
Veröffentlicht: (2023)
A Multi-Task Embedder For Retrieval Augmented LLMs
von: Zhang, Peitian, et al.
Veröffentlicht: (2023)
von: Zhang, Peitian, et al.
Veröffentlicht: (2023)
Does RAG Really Perform Bad For Long-Context Processing?
von: Luo, Kun, et al.
Veröffentlicht: (2025)
von: Luo, Kun, et al.
Veröffentlicht: (2025)
BGE Landmark Embedding: A Chunking-Free Embedding Method For Retrieval Augmented Long-Context Large Language Models
von: Luo, Kun, et al.
Veröffentlicht: (2024)
von: Luo, Kun, et al.
Veröffentlicht: (2024)
Say More with Less: Understanding Prompt Learning Behaviors through Gist Compression
von: Li, Xinze, et al.
Veröffentlicht: (2024)
von: Li, Xinze, et al.
Veröffentlicht: (2024)
INTERS: Unlocking the Power of Large Language Models in Search with Instruction Tuning
von: Zhu, Yutao, et al.
Veröffentlicht: (2024)
von: Zhu, Yutao, et al.
Veröffentlicht: (2024)
COMI: Coarse-to-fine Context Compression via Marginal Information Gain
von: Tang, Jiwei, et al.
Veröffentlicht: (2026)
von: Tang, Jiwei, et al.
Veröffentlicht: (2026)
From Matching to Generation: A Survey on Generative Information Retrieval
von: Li, Xiaoxi, et al.
Veröffentlicht: (2024)
von: Li, Xiaoxi, et al.
Veröffentlicht: (2024)
Llama2Vec: Unsupervised Adaptation of Large Language Models for Dense Retrieval
von: Liu, Zheng, et al.
Veröffentlicht: (2023)
von: Liu, Zheng, et al.
Veröffentlicht: (2023)
EmoGist: Efficient In-Context Learning for Visual Emotion Understanding
von: Seoh, Ronald, et al.
Veröffentlicht: (2025)
von: Seoh, Ronald, et al.
Veröffentlicht: (2025)
Search-o1: Agentic Search-Enhanced Large Reasoning Models
von: Li, Xiaoxi, et al.
Veröffentlicht: (2025)
von: Li, Xiaoxi, et al.
Veröffentlicht: (2025)
Investigating the Effectiveness of HyperTuning via Gisting
von: Phang, Jason
Veröffentlicht: (2024)
von: Phang, Jason
Veröffentlicht: (2024)
Grounding Language Model with Chunking-Free In-Context Retrieval
von: Qian, Hongjin, et al.
Veröffentlicht: (2024)
von: Qian, Hongjin, et al.
Veröffentlicht: (2024)
Read As Human: Compressing Context via Parallelizable Close Reading and Skimming
von: Tang, Jiwei, et al.
Veröffentlicht: (2026)
von: Tang, Jiwei, et al.
Veröffentlicht: (2026)
Matryoshka Re-Ranker: A Flexible Re-Ranking Architecture With Configurable Depth and Width
von: Liu, Zheng, et al.
Veröffentlicht: (2025)
von: Liu, Zheng, et al.
Veröffentlicht: (2025)
A Human-Inspired Reading Agent with Gist Memory of Very Long Contexts
von: Lee, Kuang-Huei, et al.
Veröffentlicht: (2024)
von: Lee, Kuang-Huei, et al.
Veröffentlicht: (2024)
GMSA: Enhancing Context Compression via Group Merging and Layer Semantic Alignment
von: Tang, Jiwei, et al.
Veröffentlicht: (2025)
von: Tang, Jiwei, et al.
Veröffentlicht: (2025)
Long Context In-Context Compression by Getting to the Gist of Gisting
von: Petrov, Aleksandar, et al.
Veröffentlicht: (2025)
von: Petrov, Aleksandar, et al.
Veröffentlicht: (2025)
Understanding Privacy Risks of Embeddings Induced by Large Language Models
von: Zhu, Zhihao, et al.
Veröffentlicht: (2024)
von: Zhu, Zhihao, et al.
Veröffentlicht: (2024)
From Lengthy to Lucid: A Systematic Literature Review on NLP Techniques for Taming Long Sentences
von: Passali, Tatiana, et al.
Veröffentlicht: (2023)
von: Passali, Tatiana, et al.
Veröffentlicht: (2023)
Perception Compressor: A Training-Free Prompt Compression Framework in Long Context Scenarios
von: Tang, Jiwei, et al.
Veröffentlicht: (2024)
von: Tang, Jiwei, et al.
Veröffentlicht: (2024)
Submodular Context Partitioning and Compression for In-Context Learning
von: Zheng, Shaoyi, et al.
Veröffentlicht: (2025)
von: Zheng, Shaoyi, et al.
Veröffentlicht: (2025)
AssistRAG: Boosting the Potential of Large Language Models with an Intelligent Information Assistant
von: Zhou, Yujia, et al.
Veröffentlicht: (2024)
von: Zhou, Yujia, et al.
Veröffentlicht: (2024)
P3: Prompts Promote Prompting
von: Zhang, Xinyu, et al.
Veröffentlicht: (2025)
von: Zhang, Xinyu, et al.
Veröffentlicht: (2025)
GREAT: Guiding Query Generation with a Trie for Recommending Related Search about Video at Kuaishou
von: Shao, Ninglu, et al.
Veröffentlicht: (2025)
von: Shao, Ninglu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Long Context Compression with Activation Beacon
von: Zhang, Peitian, et al.
Veröffentlicht: (2024) -
Extending Llama-3's Context Ten-Fold Overnight
von: Zhang, Peitian, et al.
Veröffentlicht: (2024) -
Extensible Embedding: A Flexible Multipler For LLM's Context Length
von: Shao, Ninglu, et al.
Veröffentlicht: (2024) -
Flexibly Scaling Large Language Models Contexts Through Extensible Tokenization
von: Shao, Ninglu, et al.
Veröffentlicht: (2024) -
Lighter And Better: Towards Flexible Context Adaptation For Retrieval Augmented Generation
von: Liu, Zheng, et al.
Veröffentlicht: (2024)