BM25S: Orders of magnitude faster lexical search via eager sparse scoring
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Lù, Xing Han |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Computational lexical analysis of Flamenco genres
von: Rosillo-Rodes, Pablo, et al.
Veröffentlicht: (2024)
von: Rosillo-Rodes, Pablo, et al.
Veröffentlicht: (2024)
Latent Terms: Dense Retrievers Contain Trivially Extractable BM25-ready Zipfian Vocabularies
von: Clavié, Benjamin, et al.
Veröffentlicht: (2026)
von: Clavié, Benjamin, et al.
Veröffentlicht: (2026)
InsertRank: LLMs can reason over BM25 scores to Improve Listwise Reranking
von: Seetharaman, Rahul, et al.
Veröffentlicht: (2025)
von: Seetharaman, Rahul, et al.
Veröffentlicht: (2025)
Analysing Calls to Order in German Parliamentary Debates
von: Smirnova, Nina, et al.
Veröffentlicht: (2026)
von: Smirnova, Nina, et al.
Veröffentlicht: (2026)
PaperRegister: Boosting Flexible-grained Paper Search via Hierarchical Register Indexing
von: Li, Zhuoqun, et al.
Veröffentlicht: (2025)
von: Li, Zhuoqun, et al.
Veröffentlicht: (2025)
OASIS: Order-Augmented Strategy for Improved Code Search
von: Gao, Zuchen, et al.
Veröffentlicht: (2025)
von: Gao, Zuchen, et al.
Veröffentlicht: (2025)
Principled and Scalable Diversity-Aware Retrieval via Cardinality-Constrained Binary Quadratic Programming
von: Lu, Qiheng, et al.
Veröffentlicht: (2026)
von: Lu, Qiheng, et al.
Veröffentlicht: (2026)
Towards Context-Robust LLMs: A Gated Representation Fine-tuning Approach
von: Zeng, Shenglai, et al.
Veröffentlicht: (2025)
von: Zeng, Shenglai, et al.
Veröffentlicht: (2025)
From BM25 to Corrective RAG: Benchmarking Retrieval Strategies for Text-and-Table Documents
von: Akarsu, Meftun, et al.
Veröffentlicht: (2026)
von: Akarsu, Meftun, et al.
Veröffentlicht: (2026)
Query as Anchor: Scenario-Adaptive User Representation via Large Language Model
von: Yuan, Jiahao, et al.
Veröffentlicht: (2026)
von: Yuan, Jiahao, et al.
Veröffentlicht: (2026)
Exploring LLM biases to manipulate AI search overview
von: Smirnov, Roman
Veröffentlicht: (2026)
von: Smirnov, Roman
Veröffentlicht: (2026)
DELTA: Pre-train a Discriminative Encoder for Legal Case Retrieval via Structural Word Alignment
von: Li, Haitao, et al.
Veröffentlicht: (2024)
von: Li, Haitao, et al.
Veröffentlicht: (2024)
Drop your Decoder: Pre-training with Bag-of-Word Prediction for Dense Passage Retrieval
von: Ma, Guangyuan, et al.
Veröffentlicht: (2024)
von: Ma, Guangyuan, et al.
Veröffentlicht: (2024)
Survey of Query-based Text Summarization
von: Yu, Hang, et al.
Veröffentlicht: (2022)
von: Yu, Hang, et al.
Veröffentlicht: (2022)
BBK: a simpler, faster algorithm for enumerating maximal bicliques in large sparse bipartite graphs
von: Baudin, Alexis, et al.
Veröffentlicht: (2024)
von: Baudin, Alexis, et al.
Veröffentlicht: (2024)
GLTW: Joint Improved Graph Transformer and LLM via Three-Word Language for Knowledge Graph Completion
von: Luo, Kangyang, et al.
Veröffentlicht: (2025)
von: Luo, Kangyang, et al.
Veröffentlicht: (2025)
ILCiteR: Evidence-grounded Interpretable Local Citation Recommendation
von: Roy, Sayar Ghosh, et al.
Veröffentlicht: (2024)
von: Roy, Sayar Ghosh, et al.
Veröffentlicht: (2024)
Beyond Static Dialogues: Benchmarking Realistic, Heterogeneous, and Evolving Long-Term Memory
von: Zhang, Han, et al.
Veröffentlicht: (2026)
von: Zhang, Han, et al.
Veröffentlicht: (2026)
Adapting General-Purpose Embedding Models to Private Datasets Using Keyword-based Retrieval
von: Wei, Yubai, et al.
Veröffentlicht: (2025)
von: Wei, Yubai, et al.
Veröffentlicht: (2025)
Beyond True or False: Retrieval-Augmented Hierarchical Analysis of Nuanced Claims
von: Kargupta, Priyanka, et al.
Veröffentlicht: (2025)
von: Kargupta, Priyanka, et al.
Veröffentlicht: (2025)
ReaSeq: Unleashing World Knowledge via Reasoning for Sequential Modeling
von: Tang, Jiakai, et al.
Veröffentlicht: (2025)
von: Tang, Jiakai, et al.
Veröffentlicht: (2025)
Improve Dense Passage Retrieval with Entailment Tuning
von: Dai, Lu, et al.
Veröffentlicht: (2024)
von: Dai, Lu, et al.
Veröffentlicht: (2024)
Dissecting Judicial Reasoning in U.S. Copyright Damage Awards
von: Lo, Pei-Chi, et al.
Veröffentlicht: (2026)
von: Lo, Pei-Chi, et al.
Veröffentlicht: (2026)
LADER: Log-Augmented DEnse Retrieval for Biomedical Literature Search
von: Jin, Qiao, et al.
Veröffentlicht: (2023)
von: Jin, Qiao, et al.
Veröffentlicht: (2023)
Rethinking LLM-Based Recommendations: A Personalized Query-Driven Parallel Integration
von: Han, Donghee, et al.
Veröffentlicht: (2025)
von: Han, Donghee, et al.
Veröffentlicht: (2025)
Are LLMs Reliable Rankers? Rank Manipulation via Two-Stage Token Optimization
von: Xing, Tiancheng, et al.
Veröffentlicht: (2025)
von: Xing, Tiancheng, et al.
Veröffentlicht: (2025)
TabEmbed: Benchmarking and Learning Generalist Embeddings for Tabular Understanding
von: Qiang, Minjie, et al.
Veröffentlicht: (2026)
von: Qiang, Minjie, et al.
Veröffentlicht: (2026)
WebExpert: domain-aware web agents with critic-guided expert experience for high-precision search
von: Hu, Yuelin, et al.
Veröffentlicht: (2026)
von: Hu, Yuelin, et al.
Veröffentlicht: (2026)
Entity Disambiguation via Fusion Entity Decoding
von: Wang, Junxiong, et al.
Veröffentlicht: (2024)
von: Wang, Junxiong, et al.
Veröffentlicht: (2024)
Optimizing Retrieval for RAG via Reinforcement Learning
von: Zhou, Jiawei, et al.
Veröffentlicht: (2025)
von: Zhou, Jiawei, et al.
Veröffentlicht: (2025)
Harnessing PubMed User Query Logs for Post Hoc Explanations of Recommended Similar Articles
von: Shin, Ashley, et al.
Veröffentlicht: (2024)
von: Shin, Ashley, et al.
Veröffentlicht: (2024)
QuCo-RAG: Quantifying Uncertainty from the Pre-training Corpus for Dynamic Retrieval-Augmented Generation
von: Min, Dehai, et al.
Veröffentlicht: (2025)
von: Min, Dehai, et al.
Veröffentlicht: (2025)
OkraLong: A Flexible Retrieval-Augmented Framework for Long-Text Query Processing
von: Hui, Yulong, et al.
Veröffentlicht: (2025)
von: Hui, Yulong, et al.
Veröffentlicht: (2025)
DBCopilot: Natural Language Querying over Massive Databases via Schema Routing
von: Wang, Tianshu, et al.
Veröffentlicht: (2023)
von: Wang, Tianshu, et al.
Veröffentlicht: (2023)
Decomposed Opinion Summarization with Verified Aspect-Aware Modules
von: Li, Miao, et al.
Veröffentlicht: (2025)
von: Li, Miao, et al.
Veröffentlicht: (2025)
GLEN: Generative Retrieval via Lexical Index Learning
von: Lee, Sunkyung, et al.
Veröffentlicht: (2023)
von: Lee, Sunkyung, et al.
Veröffentlicht: (2023)
Revela: Dense Retriever Learning via Language Modeling
von: Cai, Fengyu, et al.
Veröffentlicht: (2025)
von: Cai, Fengyu, et al.
Veröffentlicht: (2025)
RecDCL: Dual Contrastive Learning for Recommendation
von: Zhang, Dan, et al.
Veröffentlicht: (2024)
von: Zhang, Dan, et al.
Veröffentlicht: (2024)
Enhancing Content-based Recommendation via Large Language Model
von: Xu, Wentao, et al.
Veröffentlicht: (2024)
von: Xu, Wentao, et al.
Veröffentlicht: (2024)
Unsupervised Multilingual Dense Retrieval via Generative Pseudo Labeling
von: Huang, Chao-Wei, et al.
Veröffentlicht: (2024)
von: Huang, Chao-Wei, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Computational lexical analysis of Flamenco genres
von: Rosillo-Rodes, Pablo, et al.
Veröffentlicht: (2024) -
Latent Terms: Dense Retrievers Contain Trivially Extractable BM25-ready Zipfian Vocabularies
von: Clavié, Benjamin, et al.
Veröffentlicht: (2026) -
InsertRank: LLMs can reason over BM25 scores to Improve Listwise Reranking
von: Seetharaman, Rahul, et al.
Veröffentlicht: (2025) -
Analysing Calls to Order in German Parliamentary Debates
von: Smirnova, Nina, et al.
Veröffentlicht: (2026) -
PaperRegister: Boosting Flexible-grained Paper Search via Hierarchical Register Indexing
von: Li, Zhuoqun, et al.
Veröffentlicht: (2025)