Gespeichert in:
| Hauptverfasser: | Abdallah, Abdelrahman, Holdcroft, Jamie, Ali, Mohammed, Jatowt, Adam |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2604.03676 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TEMPO: A Realistic Multi-Domain Benchmark for Temporal Reasoning-Intensive Retrieval
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2026)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2026)
SustainableQA: A Comprehensive Question Answering Dataset for Corporate Sustainability and EU Taxonomy Reporting
von: Ali, Mohammed, et al.
Veröffentlicht: (2025)
von: Ali, Mohammed, et al.
Veröffentlicht: (2025)
RECOR: Reasoning-focused Multi-turn Conversational Retrieval Benchmark
von: Ali, Mohammed, et al.
Veröffentlicht: (2026)
von: Ali, Mohammed, et al.
Veröffentlicht: (2026)
BracketRank: Large Language Model Document Ranking via Reasoning-based Competitive Elimination
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2026)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2026)
How Good are LLM-based Rerankers? An Empirical Analysis of State-of-the-Art Reranking Models
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
Negative Sampling Techniques in Information Retrieval: A Survey
von: Wischounig, Laurin, et al.
Veröffentlicht: (2026)
von: Wischounig, Laurin, et al.
Veröffentlicht: (2026)
RankArena: A Unified Platform for Evaluating Retrieval, Reranking and RAG with Human and LLM Feedback
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
Rankify: A Comprehensive Python Toolkit for Retrieval, Re-Ranking, and Retrieval-Augmented Generation
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
DeAR: Dual-Stage Document Reranking with Reasoning Agents via LLM Distillation
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
A Study into Investigating Temporal Robustness of LLMs
von: Wallat, Jonas, et al.
Veröffentlicht: (2025)
von: Wallat, Jonas, et al.
Veröffentlicht: (2025)
TempRetriever: Fusion-based Temporal Dense Passage Retrieval for Time-Sensitive Questions
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
MM-BRIGHT: A Multi-Task Multimodal Benchmark for Reasoning-Intensive Retrieval
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2026)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2026)
Wrong Answers Can Also Be Useful: PlausibleQA -- A Large-Scale QA Dataset with Answer Plausibility Scores
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2025)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2025)
HintEval: A Comprehensive Framework for Hint Generation and Evaluation for Questions
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2025)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2025)
Exploring Hint Generation Approaches in Open-Domain Question Answering
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2024)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2024)
It's High Time: A Survey of Temporal Question Answering
von: Piryani, Bhawna, et al.
Veröffentlicht: (2025)
von: Piryani, Bhawna, et al.
Veröffentlicht: (2025)
HIVE: Query, Hypothesize, Verify An LLM Framework for Multimodal Reasoning-Intensive Retrieval
von: Abdalla, Mahmoud, et al.
Veröffentlicht: (2026)
von: Abdalla, Mahmoud, et al.
Veröffentlicht: (2026)
ArabicaQA: A Comprehensive Dataset for Arabic Question Answering
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024)
PARSE: An Open-Domain Reasoning Question Answering Benchmark for Persian
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2026)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2026)
LLMTemporalComparator: A Tool for Analysing Differences in Temporal Adaptations of Large Language Models
von: Fritsch, Reinhard Friedrich, et al.
Veröffentlicht: (2024)
von: Fritsch, Reinhard Friedrich, et al.
Veröffentlicht: (2024)
Analyzing the Role of Context in Forecasting with Large Language Models
von: Mutschlechner, Gerrit, et al.
Veröffentlicht: (2025)
von: Mutschlechner, Gerrit, et al.
Veröffentlicht: (2025)
Navigating Tomorrow: Reliably Assessing Large Language Models Performance on Future Event Prediction
von: Nako, Petraq, et al.
Veröffentlicht: (2025)
von: Nako, Petraq, et al.
Veröffentlicht: (2025)
The Impact of International Collaborations with Highly Publishing Countries in Computer Science
von: Espes, Alberto Gomez, et al.
Veröffentlicht: (2025)
von: Espes, Alberto Gomez, et al.
Veröffentlicht: (2025)
Enhancing Knowledge Retrieval with In-Context Learning and Semantic Search through Generative AI
von: Ghali, Mohammed-Khalil, et al.
Veröffentlicht: (2024)
von: Ghali, Mohammed-Khalil, et al.
Veröffentlicht: (2024)
Is Semantic Chunking Worth the Computational Cost?
von: Qu, Renyi, et al.
Veröffentlicht: (2024)
von: Qu, Renyi, et al.
Veröffentlicht: (2024)
Wisdom of the Crowds in Forecasting: Forecast Summarization for Supporting Future Event Prediction
von: Saha, Anisha, et al.
Veröffentlicht: (2025)
von: Saha, Anisha, et al.
Veröffentlicht: (2025)
Context Convergence Improves Answering Inferential Questions
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2026)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2026)
Question Difficulty Estimation for Large Language Models via Answer Plausibility Scoring
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2026)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2026)
WikiHint: A Human-Annotated Dataset for Hint Ranking and Generation
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2024)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2024)
Evaluating Answer Reranking Strategies in Time-sensitive Question Answering
von: Kardan, Mehmet, et al.
Veröffentlicht: (2025)
von: Kardan, Mehmet, et al.
Veröffentlicht: (2025)
MARVEL: Multimodal Adaptive Reasoning-intensiVe Expand-rerank and retrievaL
von: Kasem, Mahmoud SalahEldin, et al.
Veröffentlicht: (2026)
von: Kasem, Mahmoud SalahEldin, et al.
Veröffentlicht: (2026)
Detecting Future-related Contexts of Entity Mentions
von: Prashar, Puneet, et al.
Veröffentlicht: (2025)
von: Prashar, Puneet, et al.
Veröffentlicht: (2025)
Multi-hop Question Answering
von: Mavi, Vaibhav, et al.
Veröffentlicht: (2022)
von: Mavi, Vaibhav, et al.
Veröffentlicht: (2022)
Inferential Question Answering
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2026)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2026)
A Picture is Worth a Thousand Words? An Empirical Study of Aggregation Strategies for Visual Financial Document Retrieval
von: Lim, Ho Hung, et al.
Veröffentlicht: (2026)
von: Lim, Ho Hung, et al.
Veröffentlicht: (2026)
An Empirical Study of Position Bias in Modern Information Retrieval
von: Zeng, Ziyang, et al.
Veröffentlicht: (2025)
von: Zeng, Ziyang, et al.
Veröffentlicht: (2025)
Captions Are Worth a Thousand Words: Enhancing Product Retrieval with Pretrained Image-to-Text Models
von: Tang, Jason, et al.
Veröffentlicht: (2024)
von: Tang, Jason, et al.
Veröffentlicht: (2024)
Cost-Aware Retrieval-Augmentation Reasoning Models with Adaptive Retrieval Depth
von: Hashemi, Helia, et al.
Veröffentlicht: (2025)
von: Hashemi, Helia, et al.
Veröffentlicht: (2025)
AdversarialCoT: Single-Document Retrieval Poisoning for LLM Reasoning
von: Song, Hongru, et al.
Veröffentlicht: (2026)
von: Song, Hongru, et al.
Veröffentlicht: (2026)
Evaluating LLM-Based Mobile App Recommendations: An Empirical Study
von: Motger, Quim, et al.
Veröffentlicht: (2025)
von: Motger, Quim, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
TEMPO: A Realistic Multi-Domain Benchmark for Temporal Reasoning-Intensive Retrieval
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2026) -
SustainableQA: A Comprehensive Question Answering Dataset for Corporate Sustainability and EU Taxonomy Reporting
von: Ali, Mohammed, et al.
Veröffentlicht: (2025) -
RECOR: Reasoning-focused Multi-turn Conversational Retrieval Benchmark
von: Ali, Mohammed, et al.
Veröffentlicht: (2026) -
BracketRank: Large Language Model Document Ranking via Reasoning-based Competitive Elimination
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2026) -
How Good are LLM-based Rerankers? An Empirical Analysis of State-of-the-Art Reranking Models
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)