Guided Query Refinement: Multimodal Hybrid Retrieval with Test-Time Optimization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Uzan, Omri, Yehudai, Asaf, pony, Roi, Shnarch, Eyal, Gera, Ariel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Do These LLM Benchmarks Agree? Fixing Benchmark Evaluation with BenchBench
von: Perlitz, Yotam, et al.
Veröffentlicht: (2024)
von: Perlitz, Yotam, et al.
Veröffentlicht: (2024)
Teaching Values to Machines: Simulating Human-Like Behavior in LLMs
von: Yehudai, Asaf, et al.
Veröffentlicht: (2026)
von: Yehudai, Asaf, et al.
Veröffentlicht: (2026)
WildIFEval: Instruction Following in the Wild
von: Lior, Gili, et al.
Veröffentlicht: (2025)
von: Lior, Gili, et al.
Veröffentlicht: (2025)
The Mighty ToRR: A Benchmark for Table Reasoning and Robustness
von: Ashury-Tahan, Shir, et al.
Veröffentlicht: (2025)
von: Ashury-Tahan, Shir, et al.
Veröffentlicht: (2025)
Label-Efficient Model Selection for Text Generation
von: Ashury-Tahan, Shir, et al.
Veröffentlicht: (2024)
von: Ashury-Tahan, Shir, et al.
Veröffentlicht: (2024)
Genie: Achieving Human Parity in Content-Grounded Datasets Generation
von: Yehudai, Asaf, et al.
Veröffentlicht: (2024)
von: Yehudai, Asaf, et al.
Veröffentlicht: (2024)
A Nurse is Blue and Elephant is Rugby: Cross Domain Alignment in Large Language Models Reveal Human-like Patterns
von: Yehudai, Asaf, et al.
Veröffentlicht: (2024)
von: Yehudai, Asaf, et al.
Veröffentlicht: (2024)
Document Optimization for Black-Box Retrieval via Reinforcement Learning
von: Uzan, Omri, et al.
Veröffentlicht: (2026)
von: Uzan, Omri, et al.
Veröffentlicht: (2026)
Mediocrity is the key for LLM as a Judge Anchor Selection
von: Don-Yehiya, Shachar, et al.
Veröffentlicht: (2026)
von: Don-Yehiya, Shachar, et al.
Veröffentlicht: (2026)
JuStRank: Benchmarking LLM Judges for System Ranking
von: Gera, Ariel, et al.
Veröffentlicht: (2024)
von: Gera, Ariel, et al.
Veröffentlicht: (2024)
An Analysis of Hyper-Parameter Optimization Methods for Retrieval Augmented Generation
von: Orbach, Matan, et al.
Veröffentlicht: (2025)
von: Orbach, Matan, et al.
Veröffentlicht: (2025)
CharBench: Evaluating the Role of Tokenization in Character-Level Tasks
von: Uzan, Omri, et al.
Veröffentlicht: (2025)
von: Uzan, Omri, et al.
Veröffentlicht: (2025)
When LLMs are Unfit Use FastFit: Fast and Effective Text Classification with Many Classes
von: Yehudai, Asaf, et al.
Veröffentlicht: (2024)
von: Yehudai, Asaf, et al.
Veröffentlicht: (2024)
Efficient Benchmarking of Language Models
von: Perlitz, Yotam, et al.
Veröffentlicht: (2023)
von: Perlitz, Yotam, et al.
Veröffentlicht: (2023)
Task-Adaptive Embedding Refinement via Test-time LLM Guidance
von: Gera, Ariel, et al.
Veröffentlicht: (2026)
von: Gera, Ariel, et al.
Veröffentlicht: (2026)
Applying Intrinsic Debiasing on Downstream Tasks: Challenges and Considerations for Machine Translation
von: Iluz, Bar, et al.
Veröffentlicht: (2024)
von: Iluz, Bar, et al.
Veröffentlicht: (2024)
Agentic CLEAR: Automating Multi-Level Evaluation of LLM Agents
von: Yehudai, Asaf, et al.
Veröffentlicht: (2026)
von: Yehudai, Asaf, et al.
Veröffentlicht: (2026)
Greed is All You Need: An Evaluation of Tokenizer Inference Methods
von: Uzan, Omri, et al.
Veröffentlicht: (2024)
von: Uzan, Omri, et al.
Veröffentlicht: (2024)
Query Optimization for Parametric Knowledge Refinement in Retrieval-Augmented Large Language Models
von: Cong, Youan, et al.
Veröffentlicht: (2024)
von: Cong, Youan, et al.
Veröffentlicht: (2024)
RQ-RAG: Learning to Refine Queries for Retrieval Augmented Generation
von: Chan, Chi-Min, et al.
Veröffentlicht: (2024)
von: Chan, Chi-Min, et al.
Veröffentlicht: (2024)
Pretrained LLMs Learn Multiple Types of Uncertainty
von: Cohen, Roi, et al.
Veröffentlicht: (2025)
von: Cohen, Roi, et al.
Veröffentlicht: (2025)
CoRefine: Confidence-Guided Self-Refinement for Adaptive Test-Time Compute
von: Jin, Chen, et al.
Veröffentlicht: (2026)
von: Jin, Chen, et al.
Veröffentlicht: (2026)
Selective Self-to-Supervised Fine-Tuning for Generalization in Large Language Models
von: Gupta, Sonam, et al.
Veröffentlicht: (2025)
von: Gupta, Sonam, et al.
Veröffentlicht: (2025)
Will it Merge? On The Causes of Model Mergeability
von: Rahamim, Adir, et al.
Veröffentlicht: (2026)
von: Rahamim, Adir, et al.
Veröffentlicht: (2026)
Time to REFLECT: Can We Trust LLM Judges for Evidence-based Research Agents?
von: Wang, Leyao, et al.
Veröffentlicht: (2026)
von: Wang, Leyao, et al.
Veröffentlicht: (2026)
One Refiner to Unlock Them All: Inference-Time Reasoning Elicitation via Reinforcement Query Refinement
von: Zhou, Yixiao, et al.
Veröffentlicht: (2026)
von: Zhou, Yixiao, et al.
Veröffentlicht: (2026)
AdaptiVocab: Enhancing LLM Efficiency in Focused Domains through Lightweight Vocabulary Adaptation
von: Nakash, Itay, et al.
Veröffentlicht: (2025)
von: Nakash, Itay, et al.
Veröffentlicht: (2025)
CLEAR: Error Analysis via LLM-as-a-Judge Made Easy
von: Yehudai, Asaf, et al.
Veröffentlicht: (2025)
von: Yehudai, Asaf, et al.
Veröffentlicht: (2025)
Evaluating Subword Tokenization: Alien Subword Composition and OOV Generalization Challenge
von: Batsuren, Khuyagbaatar, et al.
Veröffentlicht: (2024)
von: Batsuren, Khuyagbaatar, et al.
Veröffentlicht: (2024)
Debatable Intelligence: Benchmarking LLM Judges via Debate Speech Evaluation
von: Sternlicht, Noy, et al.
Veröffentlicht: (2025)
von: Sternlicht, Noy, et al.
Veröffentlicht: (2025)
Retrieving Time-Series Differences Using Natural Language Queries
von: Dohi, Kota, et al.
Veröffentlicht: (2025)
von: Dohi, Kota, et al.
Veröffentlicht: (2025)
Motivation in Large Language Models
von: Nahum, Omer, et al.
Veröffentlicht: (2026)
von: Nahum, Omer, et al.
Veröffentlicht: (2026)
Navigating Cultural Chasms: Exploring and Unlocking the Cultural POV of Text-To-Image Models
von: Ventura, Mor, et al.
Veröffentlicht: (2023)
von: Ventura, Mor, et al.
Veröffentlicht: (2023)
Growing Pains: Extensible and Efficient LLM Benchmarking Via Fixed Parameter Calibration
von: Habba, Eliya, et al.
Veröffentlicht: (2026)
von: Habba, Eliya, et al.
Veröffentlicht: (2026)
LexiMark: Robust Watermarking via Lexical Substitutions to Enhance Membership Verification of an LLM's Textual Training Data
von: German, Eyal, et al.
Veröffentlicht: (2025)
von: German, Eyal, et al.
Veröffentlicht: (2025)
Tab-MIA: A Benchmark Dataset for Membership Inference Attacks on Tabular Data in LLMs
von: German, Eyal, et al.
Veröffentlicht: (2025)
von: German, Eyal, et al.
Veröffentlicht: (2025)
CAR: Query-Guided Confidence-Aware Reranking for Retrieval-Augmented Generation
von: Song, Zhipeng, et al.
Veröffentlicht: (2026)
von: Song, Zhipeng, et al.
Veröffentlicht: (2026)
MetaEmbed: Scaling Multimodal Retrieval at Test-Time with Flexible Late Interaction
von: Xiao, Zilin, et al.
Veröffentlicht: (2025)
von: Xiao, Zilin, et al.
Veröffentlicht: (2025)
Mixture-of-Retrieval Experts for Reasoning-Guided Multimodal Knowledge Exploitation
von: Peng, Chunyi, et al.
Veröffentlicht: (2025)
von: Peng, Chunyi, et al.
Veröffentlicht: (2025)
Selective Self-Rehearsal: A Fine-Tuning Approach to Improve Generalization in Large Language Models
von: Gupta, Sonam, et al.
Veröffentlicht: (2024)
von: Gupta, Sonam, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Do These LLM Benchmarks Agree? Fixing Benchmark Evaluation with BenchBench
von: Perlitz, Yotam, et al.
Veröffentlicht: (2024) -
Teaching Values to Machines: Simulating Human-Like Behavior in LLMs
von: Yehudai, Asaf, et al.
Veröffentlicht: (2026) -
WildIFEval: Instruction Following in the Wild
von: Lior, Gili, et al.
Veröffentlicht: (2025) -
The Mighty ToRR: A Benchmark for Table Reasoning and Robustness
von: Ashury-Tahan, Shir, et al.
Veröffentlicht: (2025) -
Label-Efficient Model Selection for Text Generation
von: Ashury-Tahan, Shir, et al.
Veröffentlicht: (2024)