Compress to Impress: Unleashing the Potential of Compressive Memory in Real-World Long-Term Conversations
Fuente:
arXiv
Salvato in:
| Autori principali: | Chen, Nuo, Li, Hongguang, Huang, Juhua, Wang, Baoyuan, Li, Jia |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Compress to Impress: Efficient LLM Adaptation Using a Single Gradient Step on 100 Samples
di: Sreeram, Shiva, et al.
Pubblicazione: (2025)
di: Sreeram, Shiva, et al.
Pubblicazione: (2025)
R1-Compress: Long Chain-of-Thought Compression via Chunk Compression and Search
di: Wang, Yibo, et al.
Pubblicazione: (2025)
di: Wang, Yibo, et al.
Pubblicazione: (2025)
DAC: A Dynamic Attention-aware Approach for Task-Agnostic Prompt Compression
di: Zhao, Yi, et al.
Pubblicazione: (2025)
di: Zhao, Yi, et al.
Pubblicazione: (2025)
REALTALK: A 21-Day Real-World Dataset for Long-Term Conversation
di: Lee, Dong-Ho, et al.
Pubblicazione: (2025)
di: Lee, Dong-Ho, et al.
Pubblicazione: (2025)
HyperMem: Hypergraph Memory for Long-Term Conversations
di: Yue, Juwei, et al.
Pubblicazione: (2026)
di: Yue, Juwei, et al.
Pubblicazione: (2026)
From Good to Great: Improving Math Reasoning with Tool-Augmented Interleaf Prompting
di: Chen, Nuo, et al.
Pubblicazione: (2023)
di: Chen, Nuo, et al.
Pubblicazione: (2023)
XQuant: Achieving Ultra-Low Bit KV Cache Quantization with Cross-Layer Compression
di: Yang, Haoqi, et al.
Pubblicazione: (2025)
di: Yang, Haoqi, et al.
Pubblicazione: (2025)
From Single to Multi-Granularity: Toward Long-Term Memory Association and Selection of Conversational Agents
di: Xu, Derong, et al.
Pubblicazione: (2025)
di: Xu, Derong, et al.
Pubblicazione: (2025)
NestedKV: Nested Memory Routing for Long-Context KV Cache Compression
di: Chen, Hong, et al.
Pubblicazione: (2026)
di: Chen, Hong, et al.
Pubblicazione: (2026)
Dynamic Long Context Reasoning over Compressed Memory via End-to-End Reinforcement Learning
di: Chen, Zhuoen, et al.
Pubblicazione: (2026)
di: Chen, Zhuoen, et al.
Pubblicazione: (2026)
EviMem: Evidence-Gap-Driven Iterative Retrieval for Long-Term Conversational Memory
di: Li, Yuyang, et al.
Pubblicazione: (2026)
di: Li, Yuyang, et al.
Pubblicazione: (2026)
Evaluating Zero-Shot Long-Context LLM Compression
di: Wang, Chenyu, et al.
Pubblicazione: (2024)
di: Wang, Chenyu, et al.
Pubblicazione: (2024)
MemRouter: Memory-as-Embedding Routing for Long-Term Conversational Agents
di: Hu, Tianyu, et al.
Pubblicazione: (2026)
di: Hu, Tianyu, et al.
Pubblicazione: (2026)
LongAttnComp: Cross-Family Context Compression for Long-Context Reasoning
di: Ji, Mengmeng, et al.
Pubblicazione: (2026)
di: Ji, Mengmeng, et al.
Pubblicazione: (2026)
When Compression Meets Model Compression: Memory-Efficient Double Compression for Large Language Models
di: Wang, Weilan, et al.
Pubblicazione: (2025)
di: Wang, Weilan, et al.
Pubblicazione: (2025)
A Compressive Memory-based Retrieval Approach for Event Argument Extraction
di: Liu, Wanlong, et al.
Pubblicazione: (2024)
di: Liu, Wanlong, et al.
Pubblicazione: (2024)
A Simple Yet Strong Baseline for Long-Term Conversational Memory of LLM Agents
di: Zhou, Sizhe, et al.
Pubblicazione: (2025)
di: Zhou, Sizhe, et al.
Pubblicazione: (2025)
Chronos: Temporal-Aware Conversational Agents with Structured Event Retrieval for Long-Term Memory
di: Sen, Sahil, et al.
Pubblicazione: (2026)
di: Sen, Sahil, et al.
Pubblicazione: (2026)
TA-Mem: Tool-Augmented Autonomous Memory Retrieval for LLM in Long-Term Conversational QA
di: Yuan, Mengwei, et al.
Pubblicazione: (2026)
di: Yuan, Mengwei, et al.
Pubblicazione: (2026)
BRIEF-Pro: Universal Context Compression with Short-to-Long Synthesis for Fast and Accurate Multi-Hop Reasoning
di: Gu, Jia-Chen, et al.
Pubblicazione: (2025)
di: Gu, Jia-Chen, et al.
Pubblicazione: (2025)
ZSMerge: Zero-Shot KV Cache Compression for Memory-Efficient Long-Context LLMs
di: Liu, Xin, et al.
Pubblicazione: (2025)
di: Liu, Xin, et al.
Pubblicazione: (2025)
Evaluating Very Long-Term Conversational Memory of LLM Agents
di: Maharana, Adyasha, et al.
Pubblicazione: (2024)
di: Maharana, Adyasha, et al.
Pubblicazione: (2024)
MemReader: From Passive to Active Extraction for Long-Term Agent Memory
di: Kang, Jingyi, et al.
Pubblicazione: (2026)
di: Kang, Jingyi, et al.
Pubblicazione: (2026)
LoMA: Lossless Compressed Memory Attention
di: Wang, Yumeng, et al.
Pubblicazione: (2024)
di: Wang, Yumeng, et al.
Pubblicazione: (2024)
EngramaBench: Evaluating Long-Term Conversational Memory with Structured Graph Retrieval
di: Acuna, Julian
Pubblicazione: (2026)
di: Acuna, Julian
Pubblicazione: (2026)
Mem-Gallery: Benchmarking Multimodal Long-Term Conversational Memory for MLLM Agents
di: Bei, Yuanchen, et al.
Pubblicazione: (2026)
di: Bei, Yuanchen, et al.
Pubblicazione: (2026)
HiGMem: A Hierarchical and LLM-Guided Memory System for Long-Term Conversational Agents
di: Cao, Shuqi, et al.
Pubblicazione: (2026)
di: Cao, Shuqi, et al.
Pubblicazione: (2026)
WebResearcher: Unleashing unbounded reasoning capability in Long-Horizon Agents
di: Qiao, Zile, et al.
Pubblicazione: (2025)
di: Qiao, Zile, et al.
Pubblicazione: (2025)
Agentic Memory: Learning Unified Long-Term and Short-Term Memory Management for Large Language Model Agents
di: Yu, Yi, et al.
Pubblicazione: (2026)
di: Yu, Yi, et al.
Pubblicazione: (2026)
Mnemis: Dual-Route Retrieval on Hierarchical Graphs for Long-Term LLM Memory
di: Tang, Zihao, et al.
Pubblicazione: (2026)
di: Tang, Zihao, et al.
Pubblicazione: (2026)
Compressing Long Context for Enhancing RAG with AMR-based Concept Distillation
di: Shi, Kaize, et al.
Pubblicazione: (2024)
di: Shi, Kaize, et al.
Pubblicazione: (2024)
M+: Extending MemoryLLM with Scalable Long-Term Memory
di: Wang, Yu, et al.
Pubblicazione: (2025)
di: Wang, Yu, et al.
Pubblicazione: (2025)
LongMemEval-V2: Evaluating Long-Term Agent Memory Toward Experienced Colleagues
di: Wu, Di, et al.
Pubblicazione: (2026)
di: Wu, Di, et al.
Pubblicazione: (2026)
Generation-Based and Emotion-Reflected Memory Update: Creating the KEEM Dataset for Better Long-Term Conversation
di: Kang, Jeonghyun, et al.
Pubblicazione: (2026)
di: Kang, Jeonghyun, et al.
Pubblicazione: (2026)
VTC-R1: Vision-Text Compression for Efficient Long-Context Reasoning
di: Wang, Yibo, et al.
Pubblicazione: (2026)
di: Wang, Yibo, et al.
Pubblicazione: (2026)
TriAttention: Efficient Long Reasoning with Trigonometric KV Compression
di: Mao, Weian, et al.
Pubblicazione: (2026)
di: Mao, Weian, et al.
Pubblicazione: (2026)
LoCoCo: Dropping In Convolutions for Long Context Compression
di: Cai, Ruisi, et al.
Pubblicazione: (2024)
di: Cai, Ruisi, et al.
Pubblicazione: (2024)
Dynamic Memory Compression: Retrofitting LLMs for Accelerated Inference
di: Nawrot, Piotr, et al.
Pubblicazione: (2024)
di: Nawrot, Piotr, et al.
Pubblicazione: (2024)
Breadcrumbs Reasoning: Memory-Efficient Reasoning with Compression Beacons
di: Monea, Giovanni, et al.
Pubblicazione: (2025)
di: Monea, Giovanni, et al.
Pubblicazione: (2025)
Goal-Directed Search Outperforms Goal-Agnostic Memory Compression in Long-Context Memory Tasks
di: Zheng, Yicong, et al.
Pubblicazione: (2025)
di: Zheng, Yicong, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Compress to Impress: Efficient LLM Adaptation Using a Single Gradient Step on 100 Samples
di: Sreeram, Shiva, et al.
Pubblicazione: (2025) -
R1-Compress: Long Chain-of-Thought Compression via Chunk Compression and Search
di: Wang, Yibo, et al.
Pubblicazione: (2025) -
DAC: A Dynamic Attention-aware Approach for Task-Agnostic Prompt Compression
di: Zhao, Yi, et al.
Pubblicazione: (2025) -
REALTALK: A 21-Day Real-World Dataset for Long-Term Conversation
di: Lee, Dong-Ho, et al.
Pubblicazione: (2025) -
HyperMem: Hypergraph Memory for Long-Term Conversations
di: Yue, Juwei, et al.
Pubblicazione: (2026)