From Volume to Value: Preference-Aligned Memory Construction for On-Device RAG
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Changmin, Kim, Jaemin, Gong, Taesik |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Aligning LLM Agents by Learning Latent Preference from User Edits
by: Gao, Ge, et al.
Published: (2024)
by: Gao, Ge, et al.
Published: (2024)
MacRAG: Compress, Slice, and Scale-up for Multi-Scale Adaptive Context RAG
by: Lim, Woosang, et al.
Published: (2025)
by: Lim, Woosang, et al.
Published: (2025)
ChatQA: Surpassing GPT-4 on Conversational QA and RAG
by: Liu, Zihan, et al.
Published: (2024)
by: Liu, Zihan, et al.
Published: (2024)
W-RAG: Weakly Supervised Dense Retrieval in RAG for Open-domain Question Answering
by: Nian, Jinming, et al.
Published: (2024)
by: Nian, Jinming, et al.
Published: (2024)
TableRAG: Million-Token Table Understanding with Language Models
by: Chen, Si-An, et al.
Published: (2024)
by: Chen, Si-An, et al.
Published: (2024)
RAG Playground: A Framework for Systematic Evaluation of Retrieval Strategies and Prompt Engineering in RAG Systems
by: Papadimitriou, Ioannis, et al.
Published: (2024)
by: Papadimitriou, Ioannis, et al.
Published: (2024)
NyayaRAG: Realistic Legal Judgment Prediction with RAG under the Indian Common Law System
by: Nigam, Shubham Kumar, et al.
Published: (2025)
by: Nigam, Shubham Kumar, et al.
Published: (2025)
Quantifying Document Impact in RAG-LLMs
by: Gerami, Armin, et al.
Published: (2025)
by: Gerami, Armin, et al.
Published: (2025)
The Chronicles of RAG: The Retriever, the Chunk and the Generator
by: Finardi, Paulo, et al.
Published: (2024)
by: Finardi, Paulo, et al.
Published: (2024)
CoRAG: Collaborative Retrieval-Augmented Generation
by: Muhamed, Aashiq, et al.
Published: (2025)
by: Muhamed, Aashiq, et al.
Published: (2025)
Emotional RAG LLMs: Reading Comprehension for the Open Internet
by: Reichman, Benjamin, et al.
Published: (2024)
by: Reichman, Benjamin, et al.
Published: (2024)
RankRAG: Unifying Context Ranking with Retrieval-Augmented Generation in LLMs
by: Yu, Yue, et al.
Published: (2024)
by: Yu, Yue, et al.
Published: (2024)
MUST-RAG: MUSical Text Question Answering with Retrieval Augmented Generation
by: Kwon, Daeyong, et al.
Published: (2025)
by: Kwon, Daeyong, et al.
Published: (2025)
Cluster-based Adaptive Retrieval: Dynamic Context Selection for RAG Applications
by: Xu, Yifan, et al.
Published: (2025)
by: Xu, Yifan, et al.
Published: (2025)
RAG Foundry: A Framework for Enhancing LLMs for Retrieval Augmented Generation
by: Fleischer, Daniel, et al.
Published: (2024)
by: Fleischer, Daniel, et al.
Published: (2024)
From Topic to Transition Structure: Unsupervised Concept Discovery at Corpus Scale via Predictive Associative Memory
by: Dury, Jason
Published: (2026)
by: Dury, Jason
Published: (2026)
Beyond RAG: Task-Aware KV Cache Compression for Comprehensive Knowledge Reasoning
by: Corallo, Giulio, et al.
Published: (2025)
by: Corallo, Giulio, et al.
Published: (2025)
Is Implicit Knowledge Enough for LLMs? A RAG Approach for Tree-based Structures
by: Gupte, Mihir, et al.
Published: (2025)
by: Gupte, Mihir, et al.
Published: (2025)
Holistic Utility Preference Learning for Listwise Alignment
by: Zhou, Jiacong, et al.
Published: (2024)
by: Zhou, Jiacong, et al.
Published: (2024)
Qwen Goes Brrr: Off-the-Shelf RAG for Ukrainian Multi-Domain Document Understanding
by: Bazdyrev, Anton, et al.
Published: (2026)
by: Bazdyrev, Anton, et al.
Published: (2026)
GraphRAG on Consumer Hardware: Benchmarking Local LLMs for Healthcare EHR Schema Retrieval
by: Fernandes, Peter, et al.
Published: (2026)
by: Fernandes, Peter, et al.
Published: (2026)
SPARC-RAG: Adaptive Sequential-Parallel Scaling with Context Management for Retrieval-Augmented Generation
by: Yang, Yuxin, et al.
Published: (2026)
by: Yang, Yuxin, et al.
Published: (2026)
ChatQA 2: Bridging the Gap to Proprietary LLMs in Long Context and RAG Capabilities
by: Xu, Peng, et al.
Published: (2024)
by: Xu, Peng, et al.
Published: (2024)
SmartChunk Retrieval: Query-Aware Chunk Compression with Planning for Efficient Document RAG
by: Zhang, Xuechen, et al.
Published: (2025)
by: Zhang, Xuechen, et al.
Published: (2025)
LightThinker++: From Reasoning Compression to Memory Management
by: Zhu, Yuqi, et al.
Published: (2026)
by: Zhu, Yuqi, et al.
Published: (2026)
RankPO: Preference Optimization for Job-Talent Matching
by: Zhang, Yafei, et al.
Published: (2025)
by: Zhang, Yafei, et al.
Published: (2025)
MetaGen Blended RAG: Unlocking Zero-Shot Precision for Specialized Domain Question-Answering
by: Sawarkar, Kunal, et al.
Published: (2025)
by: Sawarkar, Kunal, et al.
Published: (2025)
Case-Based Reasoning Approach for Solving Financial Question Answering
by: Kim, Yikyung, et al.
Published: (2024)
by: Kim, Yikyung, et al.
Published: (2024)
SimRAG: Self-Improving Retrieval-Augmented Generation for Adapting Large Language Models to Specialized Domains
by: Xu, Ran, et al.
Published: (2024)
by: Xu, Ran, et al.
Published: (2024)
Collab-RAG: Boosting Retrieval-Augmented Generation for Complex Question Answering via White-Box and Black-Box LLM Collaboration
by: Xu, Ran, et al.
Published: (2025)
by: Xu, Ran, et al.
Published: (2025)
A Systematic Evaluation of LLM Strategies for Mental Health Text Analysis: Fine-tuning vs. Prompt Engineering vs. RAG
by: Kermani, Arshia, et al.
Published: (2025)
by: Kermani, Arshia, et al.
Published: (2025)
Human-Inspired Memory Architecture for LLM Agents
by: Kerestecioglu, Doga, et al.
Published: (2026)
by: Kerestecioglu, Doga, et al.
Published: (2026)
General Agentic Memory Via Deep Research
by: Yan, B. Y., et al.
Published: (2025)
by: Yan, B. Y., et al.
Published: (2025)
ARL2: Aligning Retrievers for Black-box Large Language Models via Self-guided Adaptive Relevance Labeling
by: Zhang, Lingxi, et al.
Published: (2024)
by: Zhang, Lingxi, et al.
Published: (2024)
AriadneMem: Threading the Maze of Lifelong Memory for LLM Agents
by: Zhu, Wenhui, et al.
Published: (2026)
by: Zhu, Wenhui, et al.
Published: (2026)
A Framework for Leveraging Partially-Labeled Data for Product Attribute-Value Identification
by: Subhalingam, D., et al.
Published: (2024)
by: Subhalingam, D., et al.
Published: (2024)
Hindsight is 20/20: Building Agent Memory that Retains, Recalls, and Reflects
by: Latimer, Chris, et al.
Published: (2025)
by: Latimer, Chris, et al.
Published: (2025)
Learn to Memorize: Optimizing LLM-based Agents with Adaptive Memory Framework
by: Zhang, Zeyu, et al.
Published: (2025)
by: Zhang, Zeyu, et al.
Published: (2025)
Case-Grounded Evidence Verification: A Framework for Constructing Evidence-Sensitive Supervision
by: Arasteh, Soroosh Tayebi, et al.
Published: (2026)
by: Arasteh, Soroosh Tayebi, et al.
Published: (2026)
Scalable Frame-based Construction of Sociocultural NormBases for Socially-Aware Dialogues
by: Qu, Shilin, et al.
Published: (2024)
by: Qu, Shilin, et al.
Published: (2024)
Similar Items
-
Aligning LLM Agents by Learning Latent Preference from User Edits
by: Gao, Ge, et al.
Published: (2024) -
MacRAG: Compress, Slice, and Scale-up for Multi-Scale Adaptive Context RAG
by: Lim, Woosang, et al.
Published: (2025) -
ChatQA: Surpassing GPT-4 on Conversational QA and RAG
by: Liu, Zihan, et al.
Published: (2024) -
W-RAG: Weakly Supervised Dense Retrieval in RAG for Open-domain Question Answering
by: Nian, Jinming, et al.
Published: (2024) -
TableRAG: Million-Token Table Understanding with Language Models
by: Chen, Si-An, et al.
Published: (2024)