RubricRAG: Towards Interpretable and Reliable LLM Evaluation via Domain Knowledge Retrieval for Rubric Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Dhole, Kaustubh D., Agichtein, Eugene |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Generative Query Reformulation Using Ensemble Prompting, Document Fusion, and Relevance Feedback
di: Dhole, Kaustubh D., et al.
Pubblicazione: (2024)
di: Dhole, Kaustubh D., et al.
Pubblicazione: (2024)
To Retrieve or Not to Retrieve? Uncertainty Detection for Dynamic Retrieval Augmented Generation
di: Dhole, Kaustubh D.
Pubblicazione: (2025)
di: Dhole, Kaustubh D.
Pubblicazione: (2025)
VOGUE: A Multimodal Dataset for Conversational Recommendation in Fashion
di: Guo, David, et al.
Pubblicazione: (2025)
di: Guo, David, et al.
Pubblicazione: (2025)
Smart ETL and LLM-based contents classification: the European Smart Tourism Tools Observatory experience
di: Cosme, Diogo, et al.
Pubblicazione: (2024)
di: Cosme, Diogo, et al.
Pubblicazione: (2024)
InsertRank: LLMs can reason over BM25 scores to Improve Listwise Reranking
di: Seetharaman, Rahul, et al.
Pubblicazione: (2025)
di: Seetharaman, Rahul, et al.
Pubblicazione: (2025)
Towards Reliable Retrieval in RAG Systems for Large Legal Datasets
di: Reuter, Markus, et al.
Pubblicazione: (2025)
di: Reuter, Markus, et al.
Pubblicazione: (2025)
Evaluating User Experience in Conversational Recommender Systems: A Systematic Review Across Classical and LLM-Powered Approaches
di: Mahmud, Raj, et al.
Pubblicazione: (2025)
di: Mahmud, Raj, et al.
Pubblicazione: (2025)
BridgeRAG: Training-Free Bridge-Conditioned Retrieval for Multi-Hop Question Answering
di: Bacellar, Andre
Pubblicazione: (2026)
di: Bacellar, Andre
Pubblicazione: (2026)
KAUCUS: Knowledge Augmented User Simulators for Training Language Model Assistants
di: Dhole, Kaustubh D.
Pubblicazione: (2024)
di: Dhole, Kaustubh D.
Pubblicazione: (2024)
Experimentation Accelerator: Interpretable Insights and Creative Recommendations for A/B Testing with Content-Aware ranking
di: Hu, Zhengmian, et al.
Pubblicazione: (2026)
di: Hu, Zhengmian, et al.
Pubblicazione: (2026)
Into the Unknown Unknowns: Engaged Human Learning through Participation in Language Model Agent Conversations
di: Jiang, Yucheng, et al.
Pubblicazione: (2024)
di: Jiang, Yucheng, et al.
Pubblicazione: (2024)
A Systematic Framework for Enterprise Knowledge Retrieval: Leveraging LLM-Generated Metadata to Enhance RAG Systems
di: Mishra, Pranav Pushkar, et al.
Pubblicazione: (2025)
di: Mishra, Pranav Pushkar, et al.
Pubblicazione: (2025)
From latent factors to language: a user study on LLM-generated explanations for an inherently interpretable matrix-based recommender system
di: Manderlier, Maxime, et al.
Pubblicazione: (2025)
di: Manderlier, Maxime, et al.
Pubblicazione: (2025)
Irec: A Metacognitive Scaffolding for Self-Regulated Learning through Just-in-Time Insight Recall: A Conceptual Framework and System Prototype
di: Hou, Xuefei, et al.
Pubblicazione: (2025)
di: Hou, Xuefei, et al.
Pubblicazione: (2025)
AuthorityBench: Benchmarking LLM Authority Perception for Reliable Retrieval-Augmented Generation
di: Yao, Zhihui, et al.
Pubblicazione: (2026)
di: Yao, Zhihui, et al.
Pubblicazione: (2026)
HySemRAG: A Hybrid Semantic Retrieval-Augmented Generation Framework for Automated Literature Synthesis and Methodological Gap Analysis
di: Godinez, Alejandro
Pubblicazione: (2025)
di: Godinez, Alejandro
Pubblicazione: (2025)
From Citation Selection to Citation Absorption: A Measurement Framework for Generative Engine Optimization Across AI Search Platforms
di: Kai, Zhang, et al.
Pubblicazione: (2026)
di: Kai, Zhang, et al.
Pubblicazione: (2026)
Criteria-Based LLM Relevance Judgments
di: Farzi, Naghmeh, et al.
Pubblicazione: (2025)
di: Farzi, Naghmeh, et al.
Pubblicazione: (2025)
Architecture Matters More Than Scale: A Comparative Study of Retrieval and Memory Augmentation for Financial QA Under SME Compute Constraints
di: Liu, Jianan, et al.
Pubblicazione: (2026)
di: Liu, Jianan, et al.
Pubblicazione: (2026)
What Matters in LLM-Based Feature Extractor for Recommender? A Systematic Analysis of Prompts, Models, and Adaptation
di: Shi, Kainan, et al.
Pubblicazione: (2025)
di: Shi, Kainan, et al.
Pubblicazione: (2025)
MLDocRAG: Multimodal Long-Context Document Retrieval Augmented Generation
di: Zhang, Yongyue, et al.
Pubblicazione: (2026)
di: Zhang, Yongyue, et al.
Pubblicazione: (2026)
Less LLM, More Documents: Searching for Improved RAG
di: Ning, Jingjie, et al.
Pubblicazione: (2025)
di: Ning, Jingjie, et al.
Pubblicazione: (2025)
Introducing Axlerod: An LLM-based Chatbot for Assisting Independent Insurance Agents
di: Bradley, Adam, et al.
Pubblicazione: (2025)
di: Bradley, Adam, et al.
Pubblicazione: (2025)
AI-Friendly LaTeX: Using LaTeX Code as a Knowledge Source for Retrieval-Augmented Generation
di: Verhoeff, Tom
Pubblicazione: (2026)
di: Verhoeff, Tom
Pubblicazione: (2026)
Evaluating Perspectival Biases in Cross-Modal Retrieval
di: Saengsukhiran, Teerapol, et al.
Pubblicazione: (2025)
di: Saengsukhiran, Teerapol, et al.
Pubblicazione: (2025)
ConQRet: Benchmarking Fine-Grained Evaluation of Retrieval Augmented Argumentation with LLM Judges
di: Dhole, Kaustubh D., et al.
Pubblicazione: (2024)
di: Dhole, Kaustubh D., et al.
Pubblicazione: (2024)
Leveraging Large Language Models for Semantic Query Processing in a Scholarly Knowledge Graph
di: Jia, Runsong, et al.
Pubblicazione: (2024)
di: Jia, Runsong, et al.
Pubblicazione: (2024)
Knowledge-Aware Iterative Retrieval for Multi-Agent Systems
di: Song, Seyoung
Pubblicazione: (2025)
di: Song, Seyoung
Pubblicazione: (2025)
DeRAG: Black-box Adversarial Attacks on Multiple Retrieval-Augmented Generation Applications via Prompt Injection
di: Wang, Jerry, et al.
Pubblicazione: (2025)
di: Wang, Jerry, et al.
Pubblicazione: (2025)
Walk&Retrieve: Simple Yet Effective Zero-shot Retrieval-Augmented Generation via Knowledge Graph Walks
di: Böckling, Martin, et al.
Pubblicazione: (2025)
di: Böckling, Martin, et al.
Pubblicazione: (2025)
Prompt-RAG: Pioneering Vector Embedding-Free Retrieval-Augmented Generation in Niche Domains, Exemplified by Korean Medicine
di: Kang, Bongsu, et al.
Pubblicazione: (2024)
di: Kang, Bongsu, et al.
Pubblicazione: (2024)
MasterSet: A Large-Scale Benchmark for Must-Cite Citation Recommendation in the AI/ML Literature
di: Ratul, Md Toyaha Rahman, et al.
Pubblicazione: (2026)
di: Ratul, Md Toyaha Rahman, et al.
Pubblicazione: (2026)
Train Once, Use Flexibly: A Modular Framework for Multi-Aspect Neural News Recommendation
di: Iana, Andreea, et al.
Pubblicazione: (2023)
di: Iana, Andreea, et al.
Pubblicazione: (2023)
Session Context Embedding for Intent Understanding in Product Search
di: Mehrdad, Navid, et al.
Pubblicazione: (2024)
di: Mehrdad, Navid, et al.
Pubblicazione: (2024)
Scaling Multilingual Semantic Search in Uber Eats Delivery
di: Ling, Bo, et al.
Pubblicazione: (2026)
di: Ling, Bo, et al.
Pubblicazione: (2026)
Does UMBRELA Work on Other LLMs?
di: Farzi, Naghmeh, et al.
Pubblicazione: (2025)
di: Farzi, Naghmeh, et al.
Pubblicazione: (2025)
Algorithmic Trust and Compliance: Benchmarking Brand Notability for UK iGaming Entities in Generative Search Engines
di: Oruesagasti, Julen
Pubblicazione: (2026)
di: Oruesagasti, Julen
Pubblicazione: (2026)
Intent-Driven Dynamic Chunking: Segmenting Documents to Reflect Predicted Information Needs
di: Koutsiaris, Christos
Pubblicazione: (2026)
di: Koutsiaris, Christos
Pubblicazione: (2026)
Graph-GRPO: Dependency-Aware Credit Assignment for Generative E-commerce Search Relevance
di: Che, Jiarui, et al.
Pubblicazione: (2026)
di: Che, Jiarui, et al.
Pubblicazione: (2026)
Optimizing open-domain question answering with graph-based retrieval augmented generation
di: Cahoon, Joyce, et al.
Pubblicazione: (2025)
di: Cahoon, Joyce, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Generative Query Reformulation Using Ensemble Prompting, Document Fusion, and Relevance Feedback
di: Dhole, Kaustubh D., et al.
Pubblicazione: (2024) -
To Retrieve or Not to Retrieve? Uncertainty Detection for Dynamic Retrieval Augmented Generation
di: Dhole, Kaustubh D.
Pubblicazione: (2025) -
VOGUE: A Multimodal Dataset for Conversational Recommendation in Fashion
di: Guo, David, et al.
Pubblicazione: (2025) -
Smart ETL and LLM-based contents classification: the European Smart Tourism Tools Observatory experience
di: Cosme, Diogo, et al.
Pubblicazione: (2024) -
InsertRank: LLMs can reason over BM25 scores to Improve Listwise Reranking
di: Seetharaman, Rahul, et al.
Pubblicazione: (2025)