Understanding and Mitigating the Threat of Vec2Text to Dense Retrieval Systems
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhuang, Shengyao, Koopman, Bevan, Chu, Xiaoran, Zuccon, Guido |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Does Vec2Text Pose a New Corpus Poisoning Threat?
par: Zhuang, Shengyao, et autres
Publié: (2024)
par: Zhuang, Shengyao, et autres
Publié: (2024)
2D Matryoshka Training for Information Retrieval
par: Wang, Shuai, et autres
Publié: (2024)
par: Wang, Shuai, et autres
Publié: (2024)
Dense Retrieval with Continuous Explicit Feedback for Systematic Review Screening Prioritisation
par: Mao, Xinyu, et autres
Publié: (2024)
par: Mao, Xinyu, et autres
Publié: (2024)
Team IELAB at TREC Clinical Trial Track 2023: Enhancing Clinical Trial Retrieval with Neural Rankers and Large Language Models
par: Zhuang, Shengyao, et autres
Publié: (2024)
par: Zhuang, Shengyao, et autres
Publié: (2024)
Rank-R1: Enhancing Reasoning in LLM-based Document Rerankers via Reinforcement Learning
par: Zhuang, Shengyao, et autres
Publié: (2025)
par: Zhuang, Shengyao, et autres
Publié: (2025)
PromptReps: Prompting Large Language Models to Generate Dense and Sparse Representations for Zero-Shot Document Retrieval
par: Zhuang, Shengyao, et autres
Publié: (2024)
par: Zhuang, Shengyao, et autres
Publié: (2024)
Zero-shot Generative Large Language Models for Systematic Review Screening Automation
par: Wang, Shuai, et autres
Publié: (2024)
par: Wang, Shuai, et autres
Publié: (2024)
A Setwise Approach for Effective and Highly Efficient Zero-shot Ranking with Large Language Models
par: Zhuang, Shengyao, et autres
Publié: (2023)
par: Zhuang, Shengyao, et autres
Publié: (2023)
Reassessing Large Language Model Boolean Query Generation for Systematic Reviews
par: Wang, Shuai, et autres
Publié: (2025)
par: Wang, Shuai, et autres
Publié: (2025)
ReSLLM: Large Language Models are Strong Resource Selectors for Federated Search
par: Wang, Shuai, et autres
Publié: (2024)
par: Wang, Shuai, et autres
Publié: (2024)
LLM-VPRF: Large Language Model Based Vector Pseudo Relevance Feedback
par: Li, Hang, et autres
Publié: (2025)
par: Li, Hang, et autres
Publié: (2025)
Starbucks-v2: Improved Training for 2D Matryoshka Embeddings
par: Zhuang, Shengyao, et autres
Publié: (2024)
par: Zhuang, Shengyao, et autres
Publié: (2024)
FeB4RAG: Evaluating Federated Search in the Context of Retrieval Augmented Generation
par: Wang, Shuai, et autres
Publié: (2024)
par: Wang, Shuai, et autres
Publié: (2024)
Large Language Models for Stemming: Promises, Pitfalls and Failures
par: Wang, Shuai, et autres
Publié: (2024)
par: Wang, Shuai, et autres
Publié: (2024)
VISA: Retrieval Augmented Generation with Visual Source Attribution
par: Ma, Xueguang, et autres
Publié: (2024)
par: Ma, Xueguang, et autres
Publié: (2024)
Document Screenshot Retrievers are Vulnerable to Pixel Poisoning Attacks
par: Zhuang, Shengyao, et autres
Publié: (2025)
par: Zhuang, Shengyao, et autres
Publié: (2025)
An Investigation of Prompt Variations for Zero-shot LLM-based Rankers
par: Sun, Shuoqi, et autres
Publié: (2024)
par: Sun, Shuoqi, et autres
Publié: (2024)
Pseudo Relevance Feedback is Enough to Close the Gap Between Small and Large Dense Retrieval Models
par: Li, Hang, et autres
Publié: (2025)
par: Li, Hang, et autres
Publié: (2025)
Leveraging LLMs for Unsupervised Dense Retriever Ranking
par: Khramtsova, Ekaterina, et autres
Publié: (2024)
par: Khramtsova, Ekaterina, et autres
Publié: (2024)
Embark on DenseQuest: A System for Selecting the Best Dense Retriever for a Custom Collection
par: Khramtsova, Ekaterina, et autres
Publié: (2024)
par: Khramtsova, Ekaterina, et autres
Publié: (2024)
On the impact of retrieved content representations in RAG Pipelines
par: Ross, Jonathan J, et autres
Publié: (2026)
par: Ross, Jonathan J, et autres
Publié: (2026)
A Reproducibility Study of Goldilocks: Just-Right Tuning of BERT for TAR
par: Mao, Xinyu, et autres
Publié: (2024)
par: Mao, Xinyu, et autres
Publié: (2024)
Pre-training vs. Fine-tuning: A Reproducibility Study on Dense Retrieval Knowledge Acquisition
par: Yao, Zheng, et autres
Publié: (2025)
par: Yao, Zheng, et autres
Publié: (2025)
Beyond Chunk-Then-Embed: A Comprehensive Taxonomy and Evaluation of Document Chunking Strategies for Information Retrieval
par: Zhou, Yongjie, et autres
Publié: (2026)
par: Zhou, Yongjie, et autres
Publié: (2026)
TPRF: A Transformer-based Pseudo-Relevance Feedback Model for Efficient and Effective Retrieval
par: Li, Hang, et autres
Publié: (2024)
par: Li, Hang, et autres
Publié: (2024)
AutoBool: An Reinforcement-Learning trained LLM for Effective Automated Boolean Query Generation for Systematic Reviews
par: Wang, Shuai, et autres
Publié: (2025)
par: Wang, Shuai, et autres
Publié: (2025)
Can It Reach the Generator? Investigating the Survival of Prompt-Injection Attacks in Realistic RAG Settings
par: Yin, Yu, et autres
Publié: (2026)
par: Yin, Yu, et autres
Publié: (2026)
Rank-DistiLLM: Closing the Effectiveness Gap Between Cross-Encoders and LLMs for Passage Re-Ranking
par: Schlatt, Ferdinand, et autres
Publié: (2024)
par: Schlatt, Ferdinand, et autres
Publié: (2024)
Set-Encoder: Permutation-Invariant Inter-Passage Attention for Listwise Passage Re-Ranking with Cross-Encoders
par: Schlatt, Ferdinand, et autres
Publié: (2024)
par: Schlatt, Ferdinand, et autres
Publié: (2024)
Whole-Pool Setwise Reranking with Long-Context Language Models
par: Li, Hang, et autres
Publié: (2026)
par: Li, Hang, et autres
Publié: (2026)
LACONIC: Dense-Level Effectiveness for Scalable Sparse Retrieval via a Two-Phase Training Curriculum
par: Xu, Zhichao, et autres
Publié: (2026)
par: Xu, Zhichao, et autres
Publié: (2026)
Inferential Question Answering
par: Mozafari, Jamshid, et autres
Publié: (2026)
par: Mozafari, Jamshid, et autres
Publié: (2026)
DenseReviewer: A Screening Prioritisation Tool for Systematic Review based on Dense Retrieval
par: Mao, Xinyu, et autres
Publié: (2025)
par: Mao, Xinyu, et autres
Publié: (2025)
Distillation versus Contrastive Learning: How to Train Your Rerankers
par: Xu, Zhichao, et autres
Publié: (2025)
par: Xu, Zhichao, et autres
Publié: (2025)
BoolQuestions: Does Dense Retrieval Understand Boolean Logic in Language?
par: Zhang, Zongmeng, et autres
Publié: (2024)
par: Zhang, Zongmeng, et autres
Publié: (2024)
Dense Passage Retrieval: Is it Retrieving?
par: Reichman, Benjamin, et autres
Publié: (2024)
par: Reichman, Benjamin, et autres
Publié: (2024)
Evaluating Generative Ad Hoc Information Retrieval
par: Gienapp, Lukas, et autres
Publié: (2023)
par: Gienapp, Lukas, et autres
Publié: (2023)
Cohort Retrieval using Dense Passage Retrieval
par: Jadhav, Pranav
Publié: (2025)
par: Jadhav, Pranav
Publié: (2025)
Sparse and Dense Retrievers Learn Better Together: Joint Sparse-Dense Optimization for Text-Image Retrieval
par: Song, Jonghyun, et autres
Publié: (2025)
par: Song, Jonghyun, et autres
Publié: (2025)
Scaling Laws For Dense Retrieval
par: Fang, Yan, et autres
Publié: (2024)
par: Fang, Yan, et autres
Publié: (2024)
Documents similaires
-
Does Vec2Text Pose a New Corpus Poisoning Threat?
par: Zhuang, Shengyao, et autres
Publié: (2024) -
2D Matryoshka Training for Information Retrieval
par: Wang, Shuai, et autres
Publié: (2024) -
Dense Retrieval with Continuous Explicit Feedback for Systematic Review Screening Prioritisation
par: Mao, Xinyu, et autres
Publié: (2024) -
Team IELAB at TREC Clinical Trial Track 2023: Enhancing Clinical Trial Retrieval with Neural Rankers and Large Language Models
par: Zhuang, Shengyao, et autres
Publié: (2024) -
Rank-R1: Enhancing Reasoning in LLM-based Document Rerankers via Reinforcement Learning
par: Zhuang, Shengyao, et autres
Publié: (2025)