Provence: efficient and robust context pruning for retrieval-augmented generation
Fuente:
arXiv
Saved in:
| Main Authors: | Chirkova, Nadezhda, Formal, Thibault, Nikoulina, Vassilina, Clinchant, Stéphane |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BERGEN: A Benchmarking Library for Retrieval-Augmented Generation
by: Rau, David, et al.
Published: (2024)
by: Rau, David, et al.
Published: (2024)
Retrieval-augmented generation in multilingual settings
by: Chirkova, Nadezhda, et al.
Published: (2024)
by: Chirkova, Nadezhda, et al.
Published: (2024)
XProvence: Zero-Cost Multilingual Context Pruning for Retrieval-Augmented Generation
by: Mohamed, Youssef, et al.
Published: (2026)
by: Mohamed, Youssef, et al.
Published: (2026)
SPLADE-v3: New baselines for SPLADE
by: Lassance, Carlos, et al.
Published: (2024)
by: Lassance, Carlos, et al.
Published: (2024)
Key ingredients for effective zero-shot cross-lingual knowledge transfer in generative tasks
by: Chirkova, Nadezhda, et al.
Published: (2024)
by: Chirkova, Nadezhda, et al.
Published: (2024)
OSCAR: Online Soft Compression And Reranking
by: Louis, Maxime, et al.
Published: (2025)
by: Louis, Maxime, et al.
Published: (2025)
On the Challenges and Opportunities of Learned Sparse Retrieval for Code
by: Lupart, Simon, et al.
Published: (2026)
by: Lupart, Simon, et al.
Published: (2026)
Zero-shot cross-lingual transfer in instruction tuning of large language models
by: Chirkova, Nadezhda, et al.
Published: (2024)
by: Chirkova, Nadezhda, et al.
Published: (2024)
Empirical study of pretrained multilingual language models for zero-shot cross-lingual knowledge transfer in generation
by: Chirkova, Nadezhda, et al.
Published: (2023)
by: Chirkova, Nadezhda, et al.
Published: (2023)
A Thorough Comparison of Cross-Encoders and LLMs for Reranking SPLADE
by: Déjean, Hervé, et al.
Published: (2024)
by: Déjean, Hervé, et al.
Published: (2024)
Naver Labs Europe @ WSDM CUP | Multilingual Retrieval
by: Formal, Thibault, et al.
Published: (2026)
by: Formal, Thibault, et al.
Published: (2026)
SPLATE: Sparse Late Interaction Retrieval
by: Formal, Thibault, et al.
Published: (2024)
by: Formal, Thibault, et al.
Published: (2024)
Learning Retrieval Models with Sparse Autoencoders
by: Formal, Thibault, et al.
Published: (2026)
by: Formal, Thibault, et al.
Published: (2026)
Adapting Large Language Models for Multi-Domain Retrieval-Augmented-Generation
by: Misrahi, Alexandre, et al.
Published: (2025)
by: Misrahi, Alexandre, et al.
Published: (2025)
DiffLoRA: Differential Low-Rank Adapters for Large Language Models
by: Misrahi, Alexandre, et al.
Published: (2025)
by: Misrahi, Alexandre, et al.
Published: (2025)
Multi-Reranker: Maximizing performance of retrieval-augmented generation in the FinanceRAG challenge
by: Lee, Joohyun, et al.
Published: (2024)
by: Lee, Joohyun, et al.
Published: (2024)
Context Embeddings for Efficient Answer Generation in RAG
by: Rau, David, et al.
Published: (2024)
by: Rau, David, et al.
Published: (2024)
Investigating the potential of Sparse Mixtures-of-Experts for multi-domain neural machine translation
by: Chirkova, Nadezhda, et al.
Published: (2024)
by: Chirkova, Nadezhda, et al.
Published: (2024)
PISCO: Pretty Simple Compression for Retrieval-Augmented Generation
by: Louis, Maxime, et al.
Published: (2025)
by: Louis, Maxime, et al.
Published: (2025)
RCSB PDB AI Help Desk: retrieval-augmented generation for protein structure deposition support
by: Chithari, Vivek Reddy, et al.
Published: (2026)
by: Chithari, Vivek Reddy, et al.
Published: (2026)
Retrieval-Augmented LLM Agents: Learning to Learn from Experience
by: Ferraz, Thomas Palmeira, et al.
Published: (2026)
by: Ferraz, Thomas Palmeira, et al.
Published: (2026)
Reranking with Compressed Document Representation
by: Déjean, Hervé, et al.
Published: (2025)
by: Déjean, Hervé, et al.
Published: (2025)
Efficient Listwise Reranking with Compressed Document Representations
by: Déjean, Hervé, et al.
Published: (2026)
by: Déjean, Hervé, et al.
Published: (2026)
Multi-task retriever fine-tuning for domain-specific and efficient RAG
by: Béchard, Patrice, et al.
Published: (2025)
by: Béchard, Patrice, et al.
Published: (2025)
The Coverage Illusion: From Pre-retrieval Routing Failure to Post-retrieval Cascades in a Production RAG System
by: Hussain, Zafar, et al.
Published: (2026)
by: Hussain, Zafar, et al.
Published: (2026)
LLM2IR: simple unsupervised contrastive learning makes long-context LLM great retriever
by: Yang, Xiaocong
Published: (2025)
by: Yang, Xiaocong
Published: (2025)
FrenchToxicityPrompts: a Large Benchmark for Evaluating and Mitigating Toxicity in French Texts
by: Brun, Caroline, et al.
Published: (2024)
by: Brun, Caroline, et al.
Published: (2024)
Skill matching at scale: freelancer-project alignment for efficient multilingual candidate retrieval
by: Jouanneau, Warren, et al.
Published: (2024)
by: Jouanneau, Warren, et al.
Published: (2024)
Unveiling the Magic: Investigating Attention Distillation in Retrieval-augmented Generation
by: Li, Zizhong, et al.
Published: (2024)
by: Li, Zizhong, et al.
Published: (2024)
DAPFAM: A Domain-Aware Family-level Dataset to benchmark cross domain patent retrieval
by: Ayaou, Iliass, et al.
Published: (2025)
by: Ayaou, Iliass, et al.
Published: (2025)
Advancing continual lifelong learning in neural information retrieval: definition, dataset, framework, and empirical evaluation
by: Hou, Jingrui, et al.
Published: (2023)
by: Hou, Jingrui, et al.
Published: (2023)
Team LA at SCIDOCA shared task 2025: Citation Discovery via relation-based zero-shot retrieval
by: An, Trieu, et al.
Published: (2025)
by: An, Trieu, et al.
Published: (2025)
Knowledge graph enhanced retrieval-augmented generation for failure mode and effects analysis
by: Bahr, Lukas, et al.
Published: (2024)
by: Bahr, Lukas, et al.
Published: (2024)
Dr Web: a modern, query-based web data retrieval engine
by: Prifti, Ylli, et al.
Published: (2025)
by: Prifti, Ylli, et al.
Published: (2025)
Had enough of experts? Quantitative knowledge retrieval from large language models
by: Selby, David, et al.
Published: (2024)
by: Selby, David, et al.
Published: (2024)
Evaluation of retrieval-based QA on QUEST-LOFT
by: Scales, Nathan, et al.
Published: (2025)
by: Scales, Nathan, et al.
Published: (2025)
On the impact of retrieved content representations in RAG Pipelines
by: Ross, Jonathan J, et al.
Published: (2026)
by: Ross, Jonathan J, et al.
Published: (2026)
Two-Step SPLADE: Simple, Efficient and Effective Approximation of SPLADE
by: Lassance, Carlos, et al.
Published: (2024)
by: Lassance, Carlos, et al.
Published: (2024)
LITE: LLM-Impelled efficient Taxonomy Evaluation
by: Zhang, Lin, et al.
Published: (2025)
by: Zhang, Lin, et al.
Published: (2025)
Retrieval-augmented systems can be dangerous medical communicators
by: Wong, Lionel, et al.
Published: (2025)
by: Wong, Lionel, et al.
Published: (2025)
Similar Items
-
BERGEN: A Benchmarking Library for Retrieval-Augmented Generation
by: Rau, David, et al.
Published: (2024) -
Retrieval-augmented generation in multilingual settings
by: Chirkova, Nadezhda, et al.
Published: (2024) -
XProvence: Zero-Cost Multilingual Context Pruning for Retrieval-Augmented Generation
by: Mohamed, Youssef, et al.
Published: (2026) -
SPLADE-v3: New baselines for SPLADE
by: Lassance, Carlos, et al.
Published: (2024) -
Key ingredients for effective zero-shot cross-lingual knowledge transfer in generative tasks
by: Chirkova, Nadezhda, et al.
Published: (2024)