Likert or Not: LLM Absolute Relevance Judgments on Fine-Grained Ordinal Scales
Fuente:
arXiv
Saved in:
| Main Authors: | Godfrey, Charles, Nie, Ping, Ostapuk, Natalia, Ken, David, Gao, Shang, Inati, Souheil |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Criteria-Based LLM Relevance Judgments
by: Farzi, Naghmeh, et al.
Published: (2025)
by: Farzi, Naghmeh, et al.
Published: (2025)
Large Language Models for Relevance Judgment in Product Search
by: Mehrdad, Navid, et al.
Published: (2024)
by: Mehrdad, Navid, et al.
Published: (2024)
Memory Based Collaborative Filtering with Lucene
by: Gennaro, Claudio
Published: (2016)
by: Gennaro, Claudio
Published: (2016)
Expanding Relevance Judgments for Medical Case-based Retrieval Task with Multimodal LLMs
by: Pires, Catarina, et al.
Published: (2025)
by: Pires, Catarina, et al.
Published: (2025)
PerSoMed: A Large-Scale Balanced Dataset for Persian Social Media Text Classification
by: Chehreh, Isun, et al.
Published: (2026)
by: Chehreh, Isun, et al.
Published: (2026)
Timehash: Hierarchical Time Indexing for Efficient Business Hours Search
by: Kim, Jinoh, et al.
Published: (2026)
by: Kim, Jinoh, et al.
Published: (2026)
Formalized Information Needs Improve Large-Language-Model Relevance Judgments
by: Keller, Jüri, et al.
Published: (2026)
by: Keller, Jüri, et al.
Published: (2026)
Augmented Relevance Datasets with Fine-Tuned Small LLMs
by: Fitte-Rey, Quentin, et al.
Published: (2025)
by: Fitte-Rey, Quentin, et al.
Published: (2025)
Graph-GRPO: Dependency-Aware Credit Assignment for Generative E-commerce Search Relevance
by: Che, Jiarui, et al.
Published: (2026)
by: Che, Jiarui, et al.
Published: (2026)
TriAlignGR: Triangular Multitask Alignment with Multimodal Deep Interest Mining for Generative Recommendation
by: Zeng, Yangchen, et al.
Published: (2026)
by: Zeng, Yangchen, et al.
Published: (2026)
AI-Friendly LaTeX: Using LaTeX Code as a Knowledge Source for Retrieval-Augmented Generation
by: Verhoeff, Tom
Published: (2026)
by: Verhoeff, Tom
Published: (2026)
AlayaDB: The Data Foundation for Efficient and Effective Long-context LLM Inference
by: Deng, Yangshen, et al.
Published: (2025)
by: Deng, Yangshen, et al.
Published: (2025)
Scaling Multilingual Semantic Search in Uber Eats Delivery
by: Ling, Bo, et al.
Published: (2026)
by: Ling, Bo, et al.
Published: (2026)
Behavior-Aware Dual-Channel Preference Learning for Heterogeneous Sequential Recommendation
by: Xiao, Jing, et al.
Published: (2026)
by: Xiao, Jing, et al.
Published: (2026)
LLM-Assisted Relevance Assessments: When Should We Ask LLMs for Help?
by: Takehi, Rikiya, et al.
Published: (2024)
by: Takehi, Rikiya, et al.
Published: (2024)
MasterSet: A Large-Scale Benchmark for Must-Cite Citation Recommendation in the AI/ML Literature
by: Ratul, Md Toyaha Rahman, et al.
Published: (2026)
by: Ratul, Md Toyaha Rahman, et al.
Published: (2026)
Architecture Matters More Than Scale: A Comparative Study of Retrieval and Memory Augmentation for Financial QA Under SME Compute Constraints
by: Liu, Jianan, et al.
Published: (2026)
by: Liu, Jianan, et al.
Published: (2026)
Permanent Data Encoding (PDE): A Visual Language for Semantic Compression and Knowledge Preservation in 3-Character Units
by: Tsuyuki, Yoshiharu, et al.
Published: (2025)
by: Tsuyuki, Yoshiharu, et al.
Published: (2025)
Un análisis bibliométrico de la producción científica acerca del agrupamiento de trayectorias GPS
by: Reyes, Gary, et al.
Published: (2024)
by: Reyes, Gary, et al.
Published: (2024)
Diversification as Risk Minimization
by: Takehi, Rikiya, et al.
Published: (2025)
by: Takehi, Rikiya, et al.
Published: (2025)
From Citation Selection to Citation Absorption: A Measurement Framework for Generative Engine Optimization Across AI Search Platforms
by: Kai, Zhang, et al.
Published: (2026)
by: Kai, Zhang, et al.
Published: (2026)
AuthorityBench: Benchmarking LLM Authority Perception for Reliable Retrieval-Augmented Generation
by: Yao, Zhihui, et al.
Published: (2026)
by: Yao, Zhihui, et al.
Published: (2026)
Generative Query Reformulation Using Ensemble Prompting, Document Fusion, and Relevance Feedback
by: Dhole, Kaustubh D., et al.
Published: (2024)
by: Dhole, Kaustubh D., et al.
Published: (2024)
PeakNetFP: Peak-based Neural Audio Fingerprinting Robust to Extreme Time Stretching
by: Cortès-Sebastià, Guillem, et al.
Published: (2025)
by: Cortès-Sebastià, Guillem, et al.
Published: (2025)
Generative AI-Based Virtual Assistant using Retrieval-Augmented Generation: An evaluation study for bachelor projects
by: Verşebeniuc, Dumitru, et al.
Published: (2026)
by: Verşebeniuc, Dumitru, et al.
Published: (2026)
BridgeRAG: Training-Free Bridge-Conditioned Retrieval for Multi-Hop Question Answering
by: Bacellar, Andre
Published: (2026)
by: Bacellar, Andre
Published: (2026)
Train Once, Use Flexibly: A Modular Framework for Multi-Aspect Neural News Recommendation
by: Iana, Andreea, et al.
Published: (2023)
by: Iana, Andreea, et al.
Published: (2023)
Session Context Embedding for Intent Understanding in Product Search
by: Mehrdad, Navid, et al.
Published: (2024)
by: Mehrdad, Navid, et al.
Published: (2024)
Does UMBRELA Work on Other LLMs?
by: Farzi, Naghmeh, et al.
Published: (2025)
by: Farzi, Naghmeh, et al.
Published: (2025)
Algorithmic Trust and Compliance: Benchmarking Brand Notability for UK iGaming Entities in Generative Search Engines
by: Oruesagasti, Julen
Published: (2026)
by: Oruesagasti, Julen
Published: (2026)
Intent-Driven Dynamic Chunking: Segmenting Documents to Reflect Predicted Information Needs
by: Koutsiaris, Christos
Published: (2026)
by: Koutsiaris, Christos
Published: (2026)
Optimizing open-domain question answering with graph-based retrieval augmented generation
by: Cahoon, Joyce, et al.
Published: (2025)
by: Cahoon, Joyce, et al.
Published: (2025)
Relevance Filtering for Embedding-based Retrieval
by: Rossi, Nicholas, et al.
Published: (2024)
by: Rossi, Nicholas, et al.
Published: (2024)
Enhancing Relevance of Embedding-based Retrieval at Walmart
by: Lin, Juexin, et al.
Published: (2024)
by: Lin, Juexin, et al.
Published: (2024)
LLM-Evaluation Tropes: Perspectives on the Validity of LLM-Evaluations
by: Dietz, Laura, et al.
Published: (2025)
by: Dietz, Laura, et al.
Published: (2025)
VOGUE: A Multimodal Dataset for Conversational Recommendation in Fashion
by: Guo, David, et al.
Published: (2025)
by: Guo, David, et al.
Published: (2025)
Cross-Domain Keyword Extraction with Keyness Patterns
by: Zhou, Dongmei, et al.
Published: (2024)
by: Zhou, Dongmei, et al.
Published: (2024)
Less LLM, More Documents: Searching for Improved RAG
by: Ning, Jingjie, et al.
Published: (2025)
by: Ning, Jingjie, et al.
Published: (2025)
Multi-stage Large Language Model Pipelines Can Outperform GPT-4o in Relevance Assessment
by: Schnabel, Julian A., et al.
Published: (2025)
by: Schnabel, Julian A., et al.
Published: (2025)
Where Relevance Emerges: A Layer-Wise Study of Internal Attention for Zero-Shot Re-Ranking
by: Chen, Haodong, et al.
Published: (2026)
by: Chen, Haodong, et al.
Published: (2026)
Similar Items
-
Criteria-Based LLM Relevance Judgments
by: Farzi, Naghmeh, et al.
Published: (2025) -
Large Language Models for Relevance Judgment in Product Search
by: Mehrdad, Navid, et al.
Published: (2024) -
Memory Based Collaborative Filtering with Lucene
by: Gennaro, Claudio
Published: (2016) -
Expanding Relevance Judgments for Medical Case-based Retrieval Task with Multimodal LLMs
by: Pires, Catarina, et al.
Published: (2025) -
PerSoMed: A Large-Scale Balanced Dataset for Persian Social Media Text Classification
by: Chehreh, Isun, et al.
Published: (2026)