Hierarchical Retrieval: The Geometry and a Pretrain-Finetune Recipe
Fuente:
arXiv
Saved in:
| Main Authors: | You, Chong, Jayaram, Rajesh, Suresh, Ananda Theertha, Nittka, Robin, Yu, Felix, Kumar, Sanjiv |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adaptive Two-Phase Finetuning LLMs for Japanese Legal Text Retrieval
by: Trung, Quang Hoang, et al.
Published: (2024)
by: Trung, Quang Hoang, et al.
Published: (2024)
Scalable In-context Ranking with Generative Models
by: Gupta, Nilesh, et al.
Published: (2025)
by: Gupta, Nilesh, et al.
Published: (2025)
UltRAG: a Universal Simple Scalable Recipe for Knowledge Graph RAG
by: Georgiev, Dobrik, et al.
Published: (2026)
by: Georgiev, Dobrik, et al.
Published: (2026)
FGTR: Fine-Grained Multi-Table Retrieval via Hierarchical LLM Reasoning
by: Sun, Chaojie, et al.
Published: (2026)
by: Sun, Chaojie, et al.
Published: (2026)
RAG over Tables: Hierarchical Memory Index, Multi-Stage Retrieval, and Benchmarking
by: Zou, Jiaru, et al.
Published: (2025)
by: Zou, Jiaru, et al.
Published: (2025)
InstructRetro: Instruction Tuning post Retrieval-Augmented Pretraining
by: Wang, Boxin, et al.
Published: (2023)
by: Wang, Boxin, et al.
Published: (2023)
Diffusion-Pretrained Dense and Contextual Embeddings
by: Eslami, Sedigheh, et al.
Published: (2026)
by: Eslami, Sedigheh, et al.
Published: (2026)
An Integrated Data Processing Framework for Pretraining Foundation Models
by: Sun, Yiding, et al.
Published: (2024)
by: Sun, Yiding, et al.
Published: (2024)
Federated Learning for ICD Classification with Lightweight Models and Pretrained Embeddings
by: Xu, Binbin, et al.
Published: (2025)
by: Xu, Binbin, et al.
Published: (2025)
Scaling the Vocabulary of Non-autoregressive Models for Efficient Generative Retrieval
by: Valluri, Ravisri, et al.
Published: (2024)
by: Valluri, Ravisri, et al.
Published: (2024)
Arctic-Embed 2.0: Multilingual Retrieval Without Compromise
by: Yu, Puxuan, et al.
Published: (2024)
by: Yu, Puxuan, et al.
Published: (2024)
LEAF: Knowledge Distillation of Text Embedding Models with Teacher-Aligned Representations
by: Vujanic, Robin, et al.
Published: (2025)
by: Vujanic, Robin, et al.
Published: (2025)
Retrieval-Augmented Generation with Graphs (GraphRAG)
by: Han, Haoyu, et al.
Published: (2024)
by: Han, Haoyu, et al.
Published: (2024)
Delta Activations: A Representation for Finetuned Large Language Models
by: Xu, Zhiqiu, et al.
Published: (2025)
by: Xu, Zhiqiu, et al.
Published: (2025)
Sparse and Dense Retrievers Learn Better Together: Joint Sparse-Dense Optimization for Text-Image Retrieval
by: Song, Jonghyun, et al.
Published: (2025)
by: Song, Jonghyun, et al.
Published: (2025)
On the Theoretical Limitations of Embedding-Based Retrieval
by: Weller, Orion, et al.
Published: (2025)
by: Weller, Orion, et al.
Published: (2025)
MIRB: Mathematical Information Retrieval Benchmark
by: Ju, Haocheng, et al.
Published: (2025)
by: Ju, Haocheng, et al.
Published: (2025)
Large Language Model Can Be a Foundation for Hidden Rationale-Based Retrieval
by: Ji, Luo, et al.
Published: (2024)
by: Ji, Luo, et al.
Published: (2024)
Judgement Citation Retrieval using Contextual Similarity
by: Dasula, Akshat Mohan, et al.
Published: (2024)
by: Dasula, Akshat Mohan, et al.
Published: (2024)
IncDSI: Incrementally Updatable Document Retrieval
by: Kishore, Varsha, et al.
Published: (2023)
by: Kishore, Varsha, et al.
Published: (2023)
Retrieval-Enhanced Machine Learning: Synthesis and Opportunities
by: Kim, To Eun, et al.
Published: (2024)
by: Kim, To Eun, et al.
Published: (2024)
RankRAG: Unifying Context Ranking with Retrieval-Augmented Generation in LLMs
by: Yu, Yue, et al.
Published: (2024)
by: Yu, Yue, et al.
Published: (2024)
The Structure-Content Trade-off in Knowledge Graph Retrieval
by: Six, Valentin, et al.
Published: (2025)
by: Six, Valentin, et al.
Published: (2025)
Accelerating Retrieval-Augmented Language Model Serving with Speculation
by: Zhang, Zhihao, et al.
Published: (2024)
by: Zhang, Zhihao, et al.
Published: (2024)
Investigating Task Arithmetic for Zero-Shot Information Retrieval
by: Braga, Marco, et al.
Published: (2025)
by: Braga, Marco, et al.
Published: (2025)
From Topology to Retrieval: Decoding Embedding Spaces with Unified Signatures
by: Rottach, Florian, et al.
Published: (2025)
by: Rottach, Florian, et al.
Published: (2025)
Optimizing Multi-Stage Language Models for Effective Text Retrieval
by: Trung, Quang Hoang, et al.
Published: (2024)
by: Trung, Quang Hoang, et al.
Published: (2024)
QueryBuilder: Human-in-the-Loop Query Development for Information Retrieval
by: Kandula, Hemanth, et al.
Published: (2024)
by: Kandula, Hemanth, et al.
Published: (2024)
Rank1: Test-Time Compute for Reranking in Information Retrieval
by: Weller, Orion, et al.
Published: (2025)
by: Weller, Orion, et al.
Published: (2025)
Enhancing Question Answering Precision with Optimized Vector Retrieval and Instructions
by: Yang, Lixiao, et al.
Published: (2024)
by: Yang, Lixiao, et al.
Published: (2024)
mFollowIR: a Multilingual Benchmark for Instruction Following in Retrieval
by: Weller, Orion, et al.
Published: (2025)
by: Weller, Orion, et al.
Published: (2025)
FollowIR: Evaluating and Teaching Information Retrieval Models to Follow Instructions
by: Weller, Orion, et al.
Published: (2024)
by: Weller, Orion, et al.
Published: (2024)
Promptriever: Instruction-Trained Retrievers Can Be Prompted Like Language Models
by: Weller, Orion, et al.
Published: (2024)
by: Weller, Orion, et al.
Published: (2024)
Advancing Academic Knowledge Retrieval via LLM-enhanced Representation Similarity Fusion
by: Dai, Wei, et al.
Published: (2024)
by: Dai, Wei, et al.
Published: (2024)
Key Information Retrieval to Classify the Unstructured Data Content of Preferential Trade Agreements
by: Zhao, Jiahui, et al.
Published: (2024)
by: Zhao, Jiahui, et al.
Published: (2024)
Vectorizing the Trie: Efficient Constrained Decoding for LLM-based Generative Retrieval on Accelerators
by: Su, Zhengyang, et al.
Published: (2026)
by: Su, Zhengyang, et al.
Published: (2026)
LegalRAG: A Hybrid RAG System for Multilingual Legal Information Retrieval
by: Kabir, Muhammad Rafsan, et al.
Published: (2025)
by: Kabir, Muhammad Rafsan, et al.
Published: (2025)
GraphRAFT: Retrieval Augmented Fine-Tuning for Knowledge Graphs on Graph Databases
by: Clemedtson, Alfred, et al.
Published: (2025)
by: Clemedtson, Alfred, et al.
Published: (2025)
Beyond Sequential Reranking: Reranker-Guided Search Improves Reasoning Intensive Retrieval
by: Xu, Haike, et al.
Published: (2025)
by: Xu, Haike, et al.
Published: (2025)
Beyond Retrieval: Ensembling Cross-Encoders and GPT Rerankers with LLMs for Biomedical QA
by: Verma, Shashank, et al.
Published: (2025)
by: Verma, Shashank, et al.
Published: (2025)
Similar Items
-
Adaptive Two-Phase Finetuning LLMs for Japanese Legal Text Retrieval
by: Trung, Quang Hoang, et al.
Published: (2024) -
Scalable In-context Ranking with Generative Models
by: Gupta, Nilesh, et al.
Published: (2025) -
UltRAG: a Universal Simple Scalable Recipe for Knowledge Graph RAG
by: Georgiev, Dobrik, et al.
Published: (2026) -
FGTR: Fine-Grained Multi-Table Retrieval via Hierarchical LLM Reasoning
by: Sun, Chaojie, et al.
Published: (2026) -
RAG over Tables: Hierarchical Memory Index, Multi-Stage Retrieval, and Benchmarking
by: Zou, Jiaru, et al.
Published: (2025)