Conventional Contrastive Learning Often Falls Short: Improving Dense Retrieval with Cross-Encoder Listwise Distillation and Synthetic Data
Fuente:
arXiv
Saved in:
| Main Authors: | Tamber, Manveer Singh, Kazi, Suleman, Sourabh, Vivek, Lin, Jimmy |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Teaching Dense Retrieval Models to Specialize with Listwise Distillation and LLM Data Augmentation
by: Tamber, Manveer Singh, et al.
Published: (2025)
by: Tamber, Manveer Singh, et al.
Published: (2025)
Illusions of Relevance: Arbitrary Content Injection Attacks Deceive Retrievers, Rerankers, and LLM Judges
by: Tamber, Manveer Singh, et al.
Published: (2025)
by: Tamber, Manveer Singh, et al.
Published: (2025)
Can't Hide Behind the API: Stealing Black-Box Commercial Embedding Models
by: Tamber, Manveer Singh, et al.
Published: (2024)
by: Tamber, Manveer Singh, et al.
Published: (2024)
Unifying Adversarial Robustness and Training Across Text Scoring Models
by: Tamber, Manveer Singh, et al.
Published: (2026)
by: Tamber, Manveer Singh, et al.
Published: (2026)
Operational Advice for Dense and Sparse Retrievers: HNSW, Flat, or Inverted Indexes?
by: Lin, Jimmy
Published: (2024)
by: Lin, Jimmy
Published: (2024)
Set-Encoder: Permutation-Invariant Inter-Passage Attention for Listwise Passage Re-Ranking with Cross-Encoders
by: Schlatt, Ferdinand, et al.
Published: (2024)
by: Schlatt, Ferdinand, et al.
Published: (2024)
Negative Data Mining for Contrastive Learning in Dense Retrieval at IKEA.com
by: Agapaki, Eva, et al.
Published: (2026)
by: Agapaki, Eva, et al.
Published: (2026)
ListT5: Listwise Reranking with Fusion-in-Decoder Improves Zero-shot Retrieval
by: Yoon, Soyoung, et al.
Published: (2024)
by: Yoon, Soyoung, et al.
Published: (2024)
An Early FIRST Reproduction and Improvements to Single-Token Decoding for Fast Listwise Reranking
by: Chen, Zijian, et al.
Published: (2024)
by: Chen, Zijian, et al.
Published: (2024)
Translate-Distill: Learning Cross-Language Dense Retrieval by Translation and Distillation
by: Yang, Eugene, et al.
Published: (2024)
by: Yang, Eugene, et al.
Published: (2024)
Study on LLMs for Promptagator-Style Dense Retriever Training
by: Gwon, Daniel, et al.
Published: (2025)
by: Gwon, Daniel, et al.
Published: (2025)
CroPS: Improving Dense Retrieval with Cross-Perspective Positive Samples in Short-Video Search
by: Xie, Ao, et al.
Published: (2025)
by: Xie, Ao, et al.
Published: (2025)
Reproducing and Comparing Distillation Techniques for Cross-Encoders
by: Morand, Victor, et al.
Published: (2026)
by: Morand, Victor, et al.
Published: (2026)
Can Query Expansion Improve Generalization of Strong Cross-Encoder Rankers?
by: Li, Minghan, et al.
Published: (2023)
by: Li, Minghan, et al.
Published: (2023)
Improving the Robustness of Dense Retrievers Against Typos via Multi-Positive Contrastive Learning
by: Sidiropoulos, Georgios, et al.
Published: (2024)
by: Sidiropoulos, Georgios, et al.
Published: (2024)
LACONIC: Dense-Level Effectiveness for Scalable Sparse Retrieval via a Two-Phase Training Curriculum
by: Xu, Zhichao, et al.
Published: (2026)
by: Xu, Zhichao, et al.
Published: (2026)
Don't Retrieve, Generate: Prompting LLMs for Synthetic Training Data in Dense Retrieval
by: Sinha, Aarush
Published: (2025)
by: Sinha, Aarush
Published: (2025)
Listwise Generative Retrieval Models via a Sequential Learning Process
by: Tang, Yubao, et al.
Published: (2024)
by: Tang, Yubao, et al.
Published: (2024)
FIRST: Faster Improved Listwise Reranking with Single Token Decoding
by: Reddy, Revanth Gangi, et al.
Published: (2024)
by: Reddy, Revanth Gangi, et al.
Published: (2024)
PromptReps: Prompting Large Language Models to Generate Dense and Sparse Representations for Zero-Shot Document Retrieval
by: Zhuang, Shengyao, et al.
Published: (2024)
by: Zhuang, Shengyao, et al.
Published: (2024)
PairDistill: Pairwise Relevance Distillation for Dense Retrieval
by: Huang, Chao-Wei, et al.
Published: (2024)
by: Huang, Chao-Wei, et al.
Published: (2024)
Guiding Retrieval using LLM-based Listwise Rankers
by: Rathee, Mandeep, et al.
Published: (2025)
by: Rathee, Mandeep, et al.
Published: (2025)
SyNeg: LLM-Driven Synthetic Hard-Negatives for Dense Retrieval
by: Li, Xiaopeng, et al.
Published: (2024)
by: Li, Xiaopeng, et al.
Published: (2024)
RankGR: Rank-Enhanced Generative Retrieval with Listwise Direct Preference Optimization in Recommendation
by: Fu, Kairui, et al.
Published: (2026)
by: Fu, Kairui, et al.
Published: (2026)
On Listwise Reranking for Corpus Feedback
by: Yoon, Soyoung, et al.
Published: (2025)
by: Yoon, Soyoung, et al.
Published: (2025)
Modeling Sequential Sentence Relation to Improve Cross-lingual Dense Retrieval
by: Zhang, Shunyu, et al.
Published: (2023)
by: Zhang, Shunyu, et al.
Published: (2023)
Rank-K: Test-Time Reasoning for Listwise Reranking
by: Yang, Eugene, et al.
Published: (2025)
by: Yang, Eugene, et al.
Published: (2025)
Contextual Dual Learning Algorithm with Listwise Distillation for Unbiased Learning to Rank
by: Yu, Lulu, et al.
Published: (2024)
by: Yu, Lulu, et al.
Published: (2024)
Improving Dense Passage Retrieval with Multiple Positive Passages
by: Chang, Shuai
Published: (2025)
by: Chang, Shuai
Published: (2025)
Beyond Hard Negatives: The Importance of Score Distribution in Knowledge Distillation for Dense Retrieval
by: Jang, Youngjoon, et al.
Published: (2026)
by: Jang, Youngjoon, et al.
Published: (2026)
DRAMA: Diverse Augmentation from Large Language Models to Smaller Dense Retrievers
by: Ma, Xueguang, et al.
Published: (2025)
by: Ma, Xueguang, et al.
Published: (2025)
ListConRanker: A Contrastive Text Reranker with Listwise Encoding
by: Liu, Junlong, et al.
Published: (2025)
by: Liu, Junlong, et al.
Published: (2025)
Options-Aware Dense Retrieval for Multiple-Choice query Answering
by: Singh, Manish, et al.
Published: (2025)
by: Singh, Manish, et al.
Published: (2025)
ECLIPSE: Contrastive Dimension Importance Estimation with Pseudo-Irrelevance Feedback for Dense Retrieval
by: D'Erasmo, Giulio, et al.
Published: (2024)
by: D'Erasmo, Giulio, et al.
Published: (2024)
Boosting Data Utilization for Multilingual Dense Retrieval
by: Huang, Chao, et al.
Published: (2025)
by: Huang, Chao, et al.
Published: (2025)
A Survey of Model Architectures in Information Retrieval
by: Xu, Zhichao, et al.
Published: (2025)
by: Xu, Zhichao, et al.
Published: (2025)
Leveraging LLMs for Synthesizing Training Data Across Many Languages in Multilingual Dense Retrieval
by: Thakur, Nandan, et al.
Published: (2023)
by: Thakur, Nandan, et al.
Published: (2023)
Shallow Cross-Encoders for Low-Latency Retrieval
by: Petrov, Aleksandr V., et al.
Published: (2024)
by: Petrov, Aleksandr V., et al.
Published: (2024)
BiXSE: Improving Dense Retrieval via Probabilistic Graded Relevance Distillation
by: Tsirigotis, Christos, et al.
Published: (2025)
by: Tsirigotis, Christos, et al.
Published: (2025)
Zeroshot Listwise Learning to Rank Algorithm for Recommendation
by: Wang, Hao
Published: (2024)
by: Wang, Hao
Published: (2024)
Similar Items
-
Teaching Dense Retrieval Models to Specialize with Listwise Distillation and LLM Data Augmentation
by: Tamber, Manveer Singh, et al.
Published: (2025) -
Illusions of Relevance: Arbitrary Content Injection Attacks Deceive Retrievers, Rerankers, and LLM Judges
by: Tamber, Manveer Singh, et al.
Published: (2025) -
Can't Hide Behind the API: Stealing Black-Box Commercial Embedding Models
by: Tamber, Manveer Singh, et al.
Published: (2024) -
Unifying Adversarial Robustness and Training Across Text Scoring Models
by: Tamber, Manveer Singh, et al.
Published: (2026) -
Operational Advice for Dense and Sparse Retrievers: HNSW, Flat, or Inverted Indexes?
by: Lin, Jimmy
Published: (2024)