BM25S: Orders of magnitude faster lexical search via eager sparse scoring
Fuente:
arXiv
Salvato in:
| Autore principale: | Lù, Xing Han |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Computational lexical analysis of Flamenco genres
di: Rosillo-Rodes, Pablo, et al.
Pubblicazione: (2024)
di: Rosillo-Rodes, Pablo, et al.
Pubblicazione: (2024)
Latent Terms: Dense Retrievers Contain Trivially Extractable BM25-ready Zipfian Vocabularies
di: Clavié, Benjamin, et al.
Pubblicazione: (2026)
di: Clavié, Benjamin, et al.
Pubblicazione: (2026)
InsertRank: LLMs can reason over BM25 scores to Improve Listwise Reranking
di: Seetharaman, Rahul, et al.
Pubblicazione: (2025)
di: Seetharaman, Rahul, et al.
Pubblicazione: (2025)
Analysing Calls to Order in German Parliamentary Debates
di: Smirnova, Nina, et al.
Pubblicazione: (2026)
di: Smirnova, Nina, et al.
Pubblicazione: (2026)
PaperRegister: Boosting Flexible-grained Paper Search via Hierarchical Register Indexing
di: Li, Zhuoqun, et al.
Pubblicazione: (2025)
di: Li, Zhuoqun, et al.
Pubblicazione: (2025)
OASIS: Order-Augmented Strategy for Improved Code Search
di: Gao, Zuchen, et al.
Pubblicazione: (2025)
di: Gao, Zuchen, et al.
Pubblicazione: (2025)
Principled and Scalable Diversity-Aware Retrieval via Cardinality-Constrained Binary Quadratic Programming
di: Lu, Qiheng, et al.
Pubblicazione: (2026)
di: Lu, Qiheng, et al.
Pubblicazione: (2026)
Towards Context-Robust LLMs: A Gated Representation Fine-tuning Approach
di: Zeng, Shenglai, et al.
Pubblicazione: (2025)
di: Zeng, Shenglai, et al.
Pubblicazione: (2025)
From BM25 to Corrective RAG: Benchmarking Retrieval Strategies for Text-and-Table Documents
di: Akarsu, Meftun, et al.
Pubblicazione: (2026)
di: Akarsu, Meftun, et al.
Pubblicazione: (2026)
Query as Anchor: Scenario-Adaptive User Representation via Large Language Model
di: Yuan, Jiahao, et al.
Pubblicazione: (2026)
di: Yuan, Jiahao, et al.
Pubblicazione: (2026)
Exploring LLM biases to manipulate AI search overview
di: Smirnov, Roman
Pubblicazione: (2026)
di: Smirnov, Roman
Pubblicazione: (2026)
DELTA: Pre-train a Discriminative Encoder for Legal Case Retrieval via Structural Word Alignment
di: Li, Haitao, et al.
Pubblicazione: (2024)
di: Li, Haitao, et al.
Pubblicazione: (2024)
Drop your Decoder: Pre-training with Bag-of-Word Prediction for Dense Passage Retrieval
di: Ma, Guangyuan, et al.
Pubblicazione: (2024)
di: Ma, Guangyuan, et al.
Pubblicazione: (2024)
Survey of Query-based Text Summarization
di: Yu, Hang, et al.
Pubblicazione: (2022)
di: Yu, Hang, et al.
Pubblicazione: (2022)
BBK: a simpler, faster algorithm for enumerating maximal bicliques in large sparse bipartite graphs
di: Baudin, Alexis, et al.
Pubblicazione: (2024)
di: Baudin, Alexis, et al.
Pubblicazione: (2024)
GLTW: Joint Improved Graph Transformer and LLM via Three-Word Language for Knowledge Graph Completion
di: Luo, Kangyang, et al.
Pubblicazione: (2025)
di: Luo, Kangyang, et al.
Pubblicazione: (2025)
ILCiteR: Evidence-grounded Interpretable Local Citation Recommendation
di: Roy, Sayar Ghosh, et al.
Pubblicazione: (2024)
di: Roy, Sayar Ghosh, et al.
Pubblicazione: (2024)
Beyond Static Dialogues: Benchmarking Realistic, Heterogeneous, and Evolving Long-Term Memory
di: Zhang, Han, et al.
Pubblicazione: (2026)
di: Zhang, Han, et al.
Pubblicazione: (2026)
Adapting General-Purpose Embedding Models to Private Datasets Using Keyword-based Retrieval
di: Wei, Yubai, et al.
Pubblicazione: (2025)
di: Wei, Yubai, et al.
Pubblicazione: (2025)
Beyond True or False: Retrieval-Augmented Hierarchical Analysis of Nuanced Claims
di: Kargupta, Priyanka, et al.
Pubblicazione: (2025)
di: Kargupta, Priyanka, et al.
Pubblicazione: (2025)
ReaSeq: Unleashing World Knowledge via Reasoning for Sequential Modeling
di: Tang, Jiakai, et al.
Pubblicazione: (2025)
di: Tang, Jiakai, et al.
Pubblicazione: (2025)
Improve Dense Passage Retrieval with Entailment Tuning
di: Dai, Lu, et al.
Pubblicazione: (2024)
di: Dai, Lu, et al.
Pubblicazione: (2024)
Dissecting Judicial Reasoning in U.S. Copyright Damage Awards
di: Lo, Pei-Chi, et al.
Pubblicazione: (2026)
di: Lo, Pei-Chi, et al.
Pubblicazione: (2026)
LADER: Log-Augmented DEnse Retrieval for Biomedical Literature Search
di: Jin, Qiao, et al.
Pubblicazione: (2023)
di: Jin, Qiao, et al.
Pubblicazione: (2023)
Rethinking LLM-Based Recommendations: A Personalized Query-Driven Parallel Integration
di: Han, Donghee, et al.
Pubblicazione: (2025)
di: Han, Donghee, et al.
Pubblicazione: (2025)
Are LLMs Reliable Rankers? Rank Manipulation via Two-Stage Token Optimization
di: Xing, Tiancheng, et al.
Pubblicazione: (2025)
di: Xing, Tiancheng, et al.
Pubblicazione: (2025)
TabEmbed: Benchmarking and Learning Generalist Embeddings for Tabular Understanding
di: Qiang, Minjie, et al.
Pubblicazione: (2026)
di: Qiang, Minjie, et al.
Pubblicazione: (2026)
WebExpert: domain-aware web agents with critic-guided expert experience for high-precision search
di: Hu, Yuelin, et al.
Pubblicazione: (2026)
di: Hu, Yuelin, et al.
Pubblicazione: (2026)
Entity Disambiguation via Fusion Entity Decoding
di: Wang, Junxiong, et al.
Pubblicazione: (2024)
di: Wang, Junxiong, et al.
Pubblicazione: (2024)
Optimizing Retrieval for RAG via Reinforcement Learning
di: Zhou, Jiawei, et al.
Pubblicazione: (2025)
di: Zhou, Jiawei, et al.
Pubblicazione: (2025)
Harnessing PubMed User Query Logs for Post Hoc Explanations of Recommended Similar Articles
di: Shin, Ashley, et al.
Pubblicazione: (2024)
di: Shin, Ashley, et al.
Pubblicazione: (2024)
QuCo-RAG: Quantifying Uncertainty from the Pre-training Corpus for Dynamic Retrieval-Augmented Generation
di: Min, Dehai, et al.
Pubblicazione: (2025)
di: Min, Dehai, et al.
Pubblicazione: (2025)
OkraLong: A Flexible Retrieval-Augmented Framework for Long-Text Query Processing
di: Hui, Yulong, et al.
Pubblicazione: (2025)
di: Hui, Yulong, et al.
Pubblicazione: (2025)
DBCopilot: Natural Language Querying over Massive Databases via Schema Routing
di: Wang, Tianshu, et al.
Pubblicazione: (2023)
di: Wang, Tianshu, et al.
Pubblicazione: (2023)
Decomposed Opinion Summarization with Verified Aspect-Aware Modules
di: Li, Miao, et al.
Pubblicazione: (2025)
di: Li, Miao, et al.
Pubblicazione: (2025)
GLEN: Generative Retrieval via Lexical Index Learning
di: Lee, Sunkyung, et al.
Pubblicazione: (2023)
di: Lee, Sunkyung, et al.
Pubblicazione: (2023)
Revela: Dense Retriever Learning via Language Modeling
di: Cai, Fengyu, et al.
Pubblicazione: (2025)
di: Cai, Fengyu, et al.
Pubblicazione: (2025)
RecDCL: Dual Contrastive Learning for Recommendation
di: Zhang, Dan, et al.
Pubblicazione: (2024)
di: Zhang, Dan, et al.
Pubblicazione: (2024)
Enhancing Content-based Recommendation via Large Language Model
di: Xu, Wentao, et al.
Pubblicazione: (2024)
di: Xu, Wentao, et al.
Pubblicazione: (2024)
Unsupervised Multilingual Dense Retrieval via Generative Pseudo Labeling
di: Huang, Chao-Wei, et al.
Pubblicazione: (2024)
di: Huang, Chao-Wei, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Computational lexical analysis of Flamenco genres
di: Rosillo-Rodes, Pablo, et al.
Pubblicazione: (2024) -
Latent Terms: Dense Retrievers Contain Trivially Extractable BM25-ready Zipfian Vocabularies
di: Clavié, Benjamin, et al.
Pubblicazione: (2026) -
InsertRank: LLMs can reason over BM25 scores to Improve Listwise Reranking
di: Seetharaman, Rahul, et al.
Pubblicazione: (2025) -
Analysing Calls to Order in German Parliamentary Debates
di: Smirnova, Nina, et al.
Pubblicazione: (2026) -
PaperRegister: Boosting Flexible-grained Paper Search via Hierarchical Register Indexing
di: Li, Zhuoqun, et al.
Pubblicazione: (2025)