Identifying Key Terms in Prompts for Relevance Evaluation with GPT Models
Fuente:
arXiv
Guardado en:
| Autor principal: | Choi, Jaekeol |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Identifying Key Terms in Prompts for Relevance Evaluation with GPT Models
por: Jaekeol, Choi
Publicado: (2026)
por: Jaekeol, Choi
Publicado: (2026)
Beyond Content Relevance: Evaluating Instruction Following in Retrieval Models
por: Zhou, Jianqun, et al.
Publicado: (2024)
por: Zhou, Jianqun, et al.
Publicado: (2024)
Topic-Specific Classifiers are Better Relevance Judges than Prompted LLMs
por: Gienapp, Lukas, et al.
Publicado: (2025)
por: Gienapp, Lukas, et al.
Publicado: (2025)
The Cranfield II Relevance Assessments: A Critical Evaluation
por: Harter, Stephen P.
Publicado: (1971)
por: Harter, Stephen P.
Publicado: (1971)
Limitations of Automatic Relevance Assessments with Large Language Models for Fair and Reliable Retrieval Evaluation
por: Otero, David, et al.
Publicado: (2024)
por: Otero, David, et al.
Publicado: (2024)
LLMs Can Patch Up Missing Relevance Judgments in Evaluation
por: Upadhyay, Shivani, et al.
Publicado: (2024)
por: Upadhyay, Shivani, et al.
Publicado: (2024)
An Exam-based Evaluation Approach Beyond Traditional Relevance Judgments
por: Farzi, Naghmeh, et al.
Publicado: (2024)
por: Farzi, Naghmeh, et al.
Publicado: (2024)
Joint Evaluation of Fairness and Relevance in Recommender Systems with Pareto Frontier
por: Rampisela, Theresia Veronika, et al.
Publicado: (2025)
por: Rampisela, Theresia Veronika, et al.
Publicado: (2025)
Towards Boosting LLMs-driven Relevance Modeling with Progressive Retrieved Behavior-augmented Prompting
por: Chen, Zeyuan, et al.
Publicado: (2024)
por: Chen, Zeyuan, et al.
Publicado: (2024)
Metamorphic Evaluation of ChatGPT as a Recommender System
por: Khirbat, Madhurima, et al.
Publicado: (2024)
por: Khirbat, Madhurima, et al.
Publicado: (2024)
Unleashing the Native Recommendation Potential: LLM-Based Generative Recommendation via Structured Term Identifiers
por: Zhang, Zhiyang, et al.
Publicado: (2026)
por: Zhang, Zhiyang, et al.
Publicado: (2026)
Can We Trust Recommender System Fairness Evaluation? The Role of Fairness and Relevance
por: Rampisela, Theresia Veronika, et al.
Publicado: (2024)
por: Rampisela, Theresia Veronika, et al.
Publicado: (2024)
Do LLM-judges Align with Human Relevance in Cranfield-style Recommender Evaluation?
por: Penha, Gustavo, et al.
Publicado: (2025)
por: Penha, Gustavo, et al.
Publicado: (2025)
FGR-ColBERT: Identifying Fine-Grained Relevance Tokens During Retrieval
por: Jarolím, Antonín, et al.
Publicado: (2026)
por: Jarolím, Antonín, et al.
Publicado: (2026)
Exploring Large Language Models for Relevance Judgments in Tetun
por: de Jesus, Gabriel, et al.
Publicado: (2024)
por: de Jesus, Gabriel, et al.
Publicado: (2024)
Analyzing Adversarial Attacks on Sequence-to-Sequence Relevance Models
por: Parry, Andrew, et al.
Publicado: (2024)
por: Parry, Andrew, et al.
Publicado: (2024)
Multi-stage Large Language Model Pipelines Can Outperform GPT-4o in Relevance Assessment
por: Schnabel, Julian A., et al.
Publicado: (2025)
por: Schnabel, Julian A., et al.
Publicado: (2025)
REALM: Recursive Relevance Modeling for LLM-based Document Re-Ranking
por: Wang, Pinhuan, et al.
Publicado: (2025)
por: Wang, Pinhuan, et al.
Publicado: (2025)
HCMRM: A High-Consistency Multimodal Relevance Model for Search Ads
por: Gan, Guobing, et al.
Publicado: (2025)
por: Gan, Guobing, et al.
Publicado: (2025)
LLMJudge: LLMs for Relevance Judgments
por: Rahmani, Hossein A., et al.
Publicado: (2024)
por: Rahmani, Hossein A., et al.
Publicado: (2024)
Generalized Pseudo-Relevance Feedback
por: Tu, Yiteng, et al.
Publicado: (2025)
por: Tu, Yiteng, et al.
Publicado: (2025)
A Human-AI Comparative Analysis of Prompt Sensitivity in LLM-Based Relevance Judgment
por: Arabzadeh, Negar, et al.
Publicado: (2025)
por: Arabzadeh, Negar, et al.
Publicado: (2025)
RecGPT: A Foundation Model for Sequential Recommendation
por: Jiang, Yangqin, et al.
Publicado: (2025)
por: Jiang, Yangqin, et al.
Publicado: (2025)
Beyond Relevance: Evaluate and Improve Retrievers on Perspective Awareness
por: Zhao, Xinran, et al.
Publicado: (2024)
por: Zhao, Xinran, et al.
Publicado: (2024)
Axiomatic Causal Interventions for Reverse Engineering Relevance Computation in Neural Retrieval Models
por: Chen, Catherine, et al.
Publicado: (2024)
por: Chen, Catherine, et al.
Publicado: (2024)
Consolidating Ranking and Relevance Predictions of Large Language Models through Post-Processing
por: Yan, Le, et al.
Publicado: (2024)
por: Yan, Le, et al.
Publicado: (2024)
Discovering Biases in Information Retrieval Models Using Relevance Thesaurus as Global Explanation
por: Kim, Youngwoo, et al.
Publicado: (2024)
por: Kim, Youngwoo, et al.
Publicado: (2024)
Impact of Shallow vs. Deep Relevance Judgments on BERT-based Reranking Models
por: Iturra-Bocaz, Gabriel, et al.
Publicado: (2025)
por: Iturra-Bocaz, Gabriel, et al.
Publicado: (2025)
RAG vs. GraphRAG: A Systematic Evaluation and Key Insights
por: Han, Haoyu, et al.
Publicado: (2025)
por: Han, Haoyu, et al.
Publicado: (2025)
LLM-Assisted Pseudo-Relevance Feedback
por: Otero, David, et al.
Publicado: (2026)
por: Otero, David, et al.
Publicado: (2026)
LUCid: Redefining Relevance For Lifelong Personalization
por: Okite, Chimaobi, et al.
Publicado: (2026)
por: Okite, Chimaobi, et al.
Publicado: (2026)
TPRF: A Transformer-based Pseudo-Relevance Feedback Model for Efficient and Effective Retrieval
por: Li, Hang, et al.
Publicado: (2024)
por: Li, Hang, et al.
Publicado: (2024)
MemConflict: Evaluating Long-Term Memory Systems Under Memory Conflicts
por: Tao, Zhen, et al.
Publicado: (2026)
por: Tao, Zhen, et al.
Publicado: (2026)
UQABench: Evaluating User Embedding for Prompting LLMs in Personalized Question Answering
por: Liu, Langming, et al.
Publicado: (2025)
por: Liu, Langming, et al.
Publicado: (2025)
Impact and Relevance of Cognition Journal in the Field of Cognitive Science: An Evaluation
por: Batcha, M Sadik, et al.
Publicado: (2025)
por: Batcha, M Sadik, et al.
Publicado: (2025)
Augmenting Sequential Recommendation with Balanced Relevance and Diversity
por: Dang, Yizhou, et al.
Publicado: (2024)
por: Dang, Yizhou, et al.
Publicado: (2024)
Generative Retrieval Meets Multi-Graded Relevance
por: Tang, Yubao, et al.
Publicado: (2024)
por: Tang, Yubao, et al.
Publicado: (2024)
Don't Use LLMs to Make Relevance Judgments
por: Soboroff, Ian
Publicado: (2024)
por: Soboroff, Ian
Publicado: (2024)
Is Relevance Propagated from Retriever to Generator in RAG?
por: Tian, Fangzheng, et al.
Publicado: (2025)
por: Tian, Fangzheng, et al.
Publicado: (2025)
Migrating a Job Search Relevance Function
por: Mountain, Bennett, et al.
Publicado: (2025)
por: Mountain, Bennett, et al.
Publicado: (2025)
Ejemplares similares
-
Identifying Key Terms in Prompts for Relevance Evaluation with GPT Models
por: Jaekeol, Choi
Publicado: (2026) -
Beyond Content Relevance: Evaluating Instruction Following in Retrieval Models
por: Zhou, Jianqun, et al.
Publicado: (2024) -
Topic-Specific Classifiers are Better Relevance Judges than Prompted LLMs
por: Gienapp, Lukas, et al.
Publicado: (2025) -
The Cranfield II Relevance Assessments: A Critical Evaluation
por: Harter, Stephen P.
Publicado: (1971) -
Limitations of Automatic Relevance Assessments with Large Language Models for Fair and Reliable Retrieval Evaluation
por: Otero, David, et al.
Publicado: (2024)