ExpertRAG: Efficient RAG with Mixture of Experts -- Optimizing Context Retrieval for Adaptive LLM Responses
Fuente:
arXiv
Saved in:
| Main Author: | Gumaan, Esmail |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SymRAG: Efficient Neuro-Symbolic Retrieval Through Adaptive Query Routing
by: Hakim, Safayat Bin, et al.
Published: (2025)
by: Hakim, Safayat Bin, et al.
Published: (2025)
ECoRAG: Evidentiality-guided Compression for Long Context RAG
by: Jeong, Yeonseok, et al.
Published: (2025)
by: Jeong, Yeonseok, et al.
Published: (2025)
Cluster-based Adaptive Retrieval: Dynamic Context Selection for RAG Applications
by: Xu, Yifan, et al.
Published: (2025)
by: Xu, Yifan, et al.
Published: (2025)
L-RAG: Balancing Context and Retrieval with Entropy-Based Lazy Loading
by: Voloshyn, Sergii
Published: (2026)
by: Voloshyn, Sergii
Published: (2026)
Adaptive Chunking: Optimizing Chunking-Method Selection for RAG
by: Júnior, Paulo Roberto de Moura, et al.
Published: (2026)
by: Júnior, Paulo Roberto de Moura, et al.
Published: (2026)
SlimRAG: Retrieval without Graphs via Entity-Aware Context Selection
by: Zhang, Jiale, et al.
Published: (2025)
by: Zhang, Jiale, et al.
Published: (2025)
xRAG: Extreme Context Compression for Retrieval-augmented Generation with One Token
by: Cheng, Xin, et al.
Published: (2024)
by: Cheng, Xin, et al.
Published: (2024)
Can Long-Context Language Models Subsume Retrieval, RAG, SQL, and More?
by: Lee, Jinhyuk, et al.
Published: (2024)
by: Lee, Jinhyuk, et al.
Published: (2024)
BioRAG: A RAG-LLM Framework for Biological Question Reasoning
by: Wang, Chengrui, et al.
Published: (2024)
by: Wang, Chengrui, et al.
Published: (2024)
Blended RAG: Improving RAG (Retriever-Augmented Generation) Accuracy with Semantic Search and Hybrid Query-Based Retrievers
by: Sawarkar, Kunal, et al.
Published: (2024)
by: Sawarkar, Kunal, et al.
Published: (2024)
MacRAG: Compress, Slice, and Scale-up for Multi-Scale Adaptive Context RAG
by: Lim, Woosang, et al.
Published: (2025)
by: Lim, Woosang, et al.
Published: (2025)
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning
by: Wu, Yuhang, et al.
Published: (2026)
by: Wu, Yuhang, et al.
Published: (2026)
Multi-Type Context-Aware Conversational Recommender Systems via Mixture-of-Experts
by: Zou, Jie, et al.
Published: (2025)
by: Zou, Jie, et al.
Published: (2025)
SPARC-RAG: Adaptive Sequential-Parallel Scaling with Context Management for Retrieval-Augmented Generation
by: Yang, Yuxin, et al.
Published: (2026)
by: Yang, Yuxin, et al.
Published: (2026)
FG-RAG: Enhancing Query-Focused Summarization with Context-Aware Fine-Grained Graph RAG
by: Hong, Yubin, et al.
Published: (2025)
by: Hong, Yubin, et al.
Published: (2025)
CLI-RAG: A Retrieval-Augmented Framework for Clinically Structured and Context Aware Text Generation with LLMs
by: Keerthana, Garapati, et al.
Published: (2025)
by: Keerthana, Garapati, et al.
Published: (2025)
DoTA-RAG: Dynamic of Thought Aggregation RAG
by: Ruangtanusak, Saksorn, et al.
Published: (2025)
by: Ruangtanusak, Saksorn, et al.
Published: (2025)
RouteRAG: Efficient Retrieval-Augmented Generation from Text and Graph via Reinforcement Learning
by: Guo, Yucan, et al.
Published: (2025)
by: Guo, Yucan, et al.
Published: (2025)
AdaGReS:Adaptive Greedy Context Selection via Redundancy-Aware Scoring for Token-Budgeted RAG
by: Peng, Chao, et al.
Published: (2025)
by: Peng, Chao, et al.
Published: (2025)
RankRAG: Unifying Context Ranking with Retrieval-Augmented Generation in LLMs
by: Yu, Yue, et al.
Published: (2024)
by: Yu, Yue, et al.
Published: (2024)
GFM-RAG: Graph Foundation Model for Retrieval Augmented Generation
by: Luo, Linhao, et al.
Published: (2025)
by: Luo, Linhao, et al.
Published: (2025)
RAG-based Architectures for Drug Side Effect Retrieval in LLMs
by: Nygren, Shad, et al.
Published: (2025)
by: Nygren, Shad, et al.
Published: (2025)
Agentic Retrieval-Augmented Generation: A Survey on Agentic RAG
by: Singh, Aditi, et al.
Published: (2025)
by: Singh, Aditi, et al.
Published: (2025)
Pistis-RAG: Enhancing Retrieval-Augmented Generation with Human Feedback
by: Bai, Yu, et al.
Published: (2024)
by: Bai, Yu, et al.
Published: (2024)
Training Sparse Mixture Of Experts Text Embedding Models
by: Nussbaum, Zach, et al.
Published: (2025)
by: Nussbaum, Zach, et al.
Published: (2025)
Sustainable Digitalization of Business with Multi-Agent RAG and LLM
by: Arslan, Muhammad, et al.
Published: (2025)
by: Arslan, Muhammad, et al.
Published: (2025)
LlamaRec-LKG-RAG: A Single-Pass, Learnable Knowledge Graph-RAG Framework for LLM-Based Ranking
by: Azizi, Vahid, et al.
Published: (2025)
by: Azizi, Vahid, et al.
Published: (2025)
VersionRAG: Version-Aware Retrieval-Augmented Generation for Evolving Documents
by: Huwiler, Daniel, et al.
Published: (2025)
by: Huwiler, Daniel, et al.
Published: (2025)
DeepRAG: Thinking to Retrieve Step by Step for Large Language Models
by: Guan, Xinyan, et al.
Published: (2025)
by: Guan, Xinyan, et al.
Published: (2025)
DynaRAG: Bridging Static and Dynamic Knowledge in Retrieval-Augmented Generation
by: Liang, Penghao, et al.
Published: (2026)
by: Liang, Penghao, et al.
Published: (2026)
Towards Fair RAG: On the Impact of Fair Ranking in Retrieval-Augmented Generation
by: Kim, To Eun, et al.
Published: (2024)
by: Kim, To Eun, et al.
Published: (2024)
ProRAG: Process-Supervised Reinforcement Learning for Retrieval-Augmented Generation
by: Wang, Zhao, et al.
Published: (2026)
by: Wang, Zhao, et al.
Published: (2026)
PersianRAG: A Retrieval-Augmented Generation System for Persian Language
by: Hosseini, Hossein, et al.
Published: (2024)
by: Hosseini, Hossein, et al.
Published: (2024)
C-RAG: Certified Generation Risks for Retrieval-Augmented Language Models
by: Kang, Mintong, et al.
Published: (2024)
by: Kang, Mintong, et al.
Published: (2024)
ROGRAG: A Robustly Optimized GraphRAG Framework
by: Wang, Zhefan, et al.
Published: (2025)
by: Wang, Zhefan, et al.
Published: (2025)
SemRAG: Semantic Knowledge-Augmented RAG for Improved Question-Answering
by: Zhong, Kezhen, et al.
Published: (2025)
by: Zhong, Kezhen, et al.
Published: (2025)
Modular RAG: Transforming RAG Systems into LEGO-like Reconfigurable Frameworks
by: Gao, Yunfan, et al.
Published: (2024)
by: Gao, Yunfan, et al.
Published: (2024)
Principled Context Engineering for RAG: Statistical Guarantees via Conformal Prediction
by: Chakraborty, Debashish, et al.
Published: (2025)
by: Chakraborty, Debashish, et al.
Published: (2025)
Tuning LLMs by RAG Principles: Towards LLM-native Memory
by: Wei, Jiale, et al.
Published: (2025)
by: Wei, Jiale, et al.
Published: (2025)
Advanced ingestion process powered by LLM parsing for RAG system
by: Perez, Arnau, et al.
Published: (2024)
by: Perez, Arnau, et al.
Published: (2024)
Similar Items
-
SymRAG: Efficient Neuro-Symbolic Retrieval Through Adaptive Query Routing
by: Hakim, Safayat Bin, et al.
Published: (2025) -
ECoRAG: Evidentiality-guided Compression for Long Context RAG
by: Jeong, Yeonseok, et al.
Published: (2025) -
Cluster-based Adaptive Retrieval: Dynamic Context Selection for RAG Applications
by: Xu, Yifan, et al.
Published: (2025) -
L-RAG: Balancing Context and Retrieval with Entropy-Based Lazy Loading
by: Voloshyn, Sergii
Published: (2026) -
Adaptive Chunking: Optimizing Chunking-Method Selection for RAG
by: Júnior, Paulo Roberto de Moura, et al.
Published: (2026)