Does UMBRELA Work on Other LLMs?
Fuente:
arXiv
Saved in:
| Main Authors: | Farzi, Naghmeh, Dietz, Laura |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Criteria-Based LLM Relevance Judgments
by: Farzi, Naghmeh, et al.
Published: (2025)
by: Farzi, Naghmeh, et al.
Published: (2025)
Optimizing open-domain question answering with graph-based retrieval augmented generation
by: Cahoon, Joyce, et al.
Published: (2025)
by: Cahoon, Joyce, et al.
Published: (2025)
MasterSet: A Large-Scale Benchmark for Must-Cite Citation Recommendation in the AI/ML Literature
by: Ratul, Md Toyaha Rahman, et al.
Published: (2026)
by: Ratul, Md Toyaha Rahman, et al.
Published: (2026)
BridgeRAG: Training-Free Bridge-Conditioned Retrieval for Multi-Hop Question Answering
by: Bacellar, Andre
Published: (2026)
by: Bacellar, Andre
Published: (2026)
Train Once, Use Flexibly: A Modular Framework for Multi-Aspect Neural News Recommendation
by: Iana, Andreea, et al.
Published: (2023)
by: Iana, Andreea, et al.
Published: (2023)
Session Context Embedding for Intent Understanding in Product Search
by: Mehrdad, Navid, et al.
Published: (2024)
by: Mehrdad, Navid, et al.
Published: (2024)
Scaling Multilingual Semantic Search in Uber Eats Delivery
by: Ling, Bo, et al.
Published: (2026)
by: Ling, Bo, et al.
Published: (2026)
Algorithmic Trust and Compliance: Benchmarking Brand Notability for UK iGaming Entities in Generative Search Engines
by: Oruesagasti, Julen
Published: (2026)
by: Oruesagasti, Julen
Published: (2026)
Intent-Driven Dynamic Chunking: Segmenting Documents to Reflect Predicted Information Needs
by: Koutsiaris, Christos
Published: (2026)
by: Koutsiaris, Christos
Published: (2026)
Graph-GRPO: Dependency-Aware Credit Assignment for Generative E-commerce Search Relevance
by: Che, Jiarui, et al.
Published: (2026)
by: Che, Jiarui, et al.
Published: (2026)
Architecture Matters More Than Scale: A Comparative Study of Retrieval and Memory Augmentation for Financial QA Under SME Compute Constraints
by: Liu, Jianan, et al.
Published: (2026)
by: Liu, Jianan, et al.
Published: (2026)
Best in Tau@LLMJudge: Criteria-Based Relevance Evaluation with Llama3
by: Farzi, Naghmeh, et al.
Published: (2024)
by: Farzi, Naghmeh, et al.
Published: (2024)
Augmented Relevance Datasets with Fine-Tuned Small LLMs
by: Fitte-Rey, Quentin, et al.
Published: (2025)
by: Fitte-Rey, Quentin, et al.
Published: (2025)
Expanding Relevance Judgments for Medical Case-based Retrieval Task with Multimodal LLMs
by: Pires, Catarina, et al.
Published: (2025)
by: Pires, Catarina, et al.
Published: (2025)
Pre-trained LLMs Meet Sequential Recommenders: Efficient User-Centric Knowledge Distillation
by: Severin, Nikita, et al.
Published: (2026)
by: Severin, Nikita, et al.
Published: (2026)
Harnessing multiple LLMs for Information Retrieval: A case study on Deep Learning methodologies in Biodiversity publications
by: Kommineni, Vamsi Krishna, et al.
Published: (2024)
by: Kommineni, Vamsi Krishna, et al.
Published: (2024)
RLM-on-KG: Heuristics First, LLMs When Needed: Adaptive Retrieval Control over Mention Graphs for Scattered Evidence
by: Volpini, Andrea, et al.
Published: (2026)
by: Volpini, Andrea, et al.
Published: (2026)
What Matters in LLM-Based Feature Extractor for Recommender? A Systematic Analysis of Prompts, Models, and Adaptation
by: Shi, Kainan, et al.
Published: (2025)
by: Shi, Kainan, et al.
Published: (2025)
Diversification as Risk Minimization
by: Takehi, Rikiya, et al.
Published: (2025)
by: Takehi, Rikiya, et al.
Published: (2025)
From Citation Selection to Citation Absorption: A Measurement Framework for Generative Engine Optimization Across AI Search Platforms
by: Kai, Zhang, et al.
Published: (2026)
by: Kai, Zhang, et al.
Published: (2026)
VOGUE: A Multimodal Dataset for Conversational Recommendation in Fashion
by: Guo, David, et al.
Published: (2025)
by: Guo, David, et al.
Published: (2025)
AuthorityBench: Benchmarking LLM Authority Perception for Reliable Retrieval-Augmented Generation
by: Yao, Zhihui, et al.
Published: (2026)
by: Yao, Zhihui, et al.
Published: (2026)
Reviewing the Reviewer: Graph-Enhanced LLMs for E-commerce Appeal Adjudication
by: Du, Yuchen, et al.
Published: (2026)
by: Du, Yuchen, et al.
Published: (2026)
Uncovering the Limitations of Query Performance Prediction: Failures, Insights, and Implications for Selective Query Processing
by: Chifu, Adrian-Gabriel, et al.
Published: (2025)
by: Chifu, Adrian-Gabriel, et al.
Published: (2025)
A Case Study of Balanced Query Recommendation on Wikipedia
by: Mishra, Harshit, et al.
Published: (2025)
by: Mishra, Harshit, et al.
Published: (2025)
WisPaper: Your AI Scholar Search Engine
by: Ju, Li, et al.
Published: (2025)
by: Ju, Li, et al.
Published: (2025)
Less LLM, More Documents: Searching for Improved RAG
by: Ning, Jingjie, et al.
Published: (2025)
by: Ning, Jingjie, et al.
Published: (2025)
Query-Centric Graph Retrieval Augmented Generation
by: Wu, Yaxiong, et al.
Published: (2025)
by: Wu, Yaxiong, et al.
Published: (2025)
Evaluating the Effectiveness of Large Language Models in Automated News Article Summarization
by: Houamegni, Lionel Richy Panlap, et al.
Published: (2025)
by: Houamegni, Lionel Richy Panlap, et al.
Published: (2025)
Walk&Retrieve: Simple Yet Effective Zero-shot Retrieval-Augmented Generation via Knowledge Graph Walks
by: Böckling, Martin, et al.
Published: (2025)
by: Böckling, Martin, et al.
Published: (2025)
ESGBench: A Benchmark for Explainable ESG Question Answering in Corporate Sustainability Reports
by: George, Sherine, et al.
Published: (2025)
by: George, Sherine, et al.
Published: (2025)
SGMem: Sentence Graph Memory for Long-Term Conversational Agents
by: Wu, Yaxiong, et al.
Published: (2025)
by: Wu, Yaxiong, et al.
Published: (2025)
A Systematic Framework for Enterprise Knowledge Retrieval: Leveraging LLM-Generated Metadata to Enhance RAG Systems
by: Mishra, Pranav Pushkar, et al.
Published: (2025)
by: Mishra, Pranav Pushkar, et al.
Published: (2025)
MLDocRAG: Multimodal Long-Context Document Retrieval Augmented Generation
by: Zhang, Yongyue, et al.
Published: (2026)
by: Zhang, Yongyue, et al.
Published: (2026)
Temporal Decay of Co-Citation Predictability: A 20-Year Statute Retrieval Benchmark from 396M Ukrainian Court Citations
by: Ovcharov, Volodymyr
Published: (2026)
by: Ovcharov, Volodymyr
Published: (2026)
MeVer at CheckThat! 2026: Cluster-Aware Hard-Negative Mining for Multilingual Scientific-Source Retrieval
by: Bakagianni, Juli, et al.
Published: (2026)
by: Bakagianni, Juli, et al.
Published: (2026)
Generative Query Reformulation Using Ensemble Prompting, Document Fusion, and Relevance Feedback
by: Dhole, Kaustubh D., et al.
Published: (2024)
by: Dhole, Kaustubh D., et al.
Published: (2024)
Exploring Information Retrieval Landscapes: An Investigation of a Novel Evaluation Techniques and Comparative Document Splitting Methods
by: Narimissa, Esmaeil, et al.
Published: (2024)
by: Narimissa, Esmaeil, et al.
Published: (2024)
Prompt Perturbation in Retrieval-Augmented Generation based Large Language Models
by: Hu, Zhibo, et al.
Published: (2024)
by: Hu, Zhibo, et al.
Published: (2024)
Large Language Models for Relevance Judgment in Product Search
by: Mehrdad, Navid, et al.
Published: (2024)
by: Mehrdad, Navid, et al.
Published: (2024)
Similar Items
-
Criteria-Based LLM Relevance Judgments
by: Farzi, Naghmeh, et al.
Published: (2025) -
Optimizing open-domain question answering with graph-based retrieval augmented generation
by: Cahoon, Joyce, et al.
Published: (2025) -
MasterSet: A Large-Scale Benchmark for Must-Cite Citation Recommendation in the AI/ML Literature
by: Ratul, Md Toyaha Rahman, et al.
Published: (2026) -
BridgeRAG: Training-Free Bridge-Conditioned Retrieval for Multi-Hop Question Answering
by: Bacellar, Andre
Published: (2026) -
Train Once, Use Flexibly: A Modular Framework for Multi-Aspect Neural News Recommendation
by: Iana, Andreea, et al.
Published: (2023)