The Decoy Dilemma in Online Medical Information Evaluation: A Comparative Study of Credibility Assessments by LLM and Human Judges
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Jiqun, He, Jiangen |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluating User Experience in Conversational Recommender Systems: A Systematic Review Across Classical and LLM-Powered Approaches
by: Mahmud, Raj, et al.
Published: (2025)
by: Mahmud, Raj, et al.
Published: (2025)
Structural Feature Engineering for Generative Engine Optimization: How Content Structure Shapes Citation Behavior
by: Yu, Junwei, et al.
Published: (2026)
by: Yu, Junwei, et al.
Published: (2026)
From latent factors to language: a user study on LLM-generated explanations for an inherently interpretable matrix-based recommender system
by: Manderlier, Maxime, et al.
Published: (2025)
by: Manderlier, Maxime, et al.
Published: (2025)
Architecture Matters More Than Scale: A Comparative Study of Retrieval and Memory Augmentation for Financial QA Under SME Compute Constraints
by: Liu, Jianan, et al.
Published: (2026)
by: Liu, Jianan, et al.
Published: (2026)
Irec: A Metacognitive Scaffolding for Self-Regulated Learning through Just-in-Time Insight Recall: A Conceptual Framework and System Prototype
by: Hou, Xuefei, et al.
Published: (2025)
by: Hou, Xuefei, et al.
Published: (2025)
Criteria-Based LLM Relevance Judgments
by: Farzi, Naghmeh, et al.
Published: (2025)
by: Farzi, Naghmeh, et al.
Published: (2025)
KAUCUS: Knowledge Augmented User Simulators for Training Language Model Assistants
by: Dhole, Kaustubh D.
Published: (2024)
by: Dhole, Kaustubh D.
Published: (2024)
Introducing Axlerod: An LLM-based Chatbot for Assisting Independent Insurance Agents
by: Bradley, Adam, et al.
Published: (2025)
by: Bradley, Adam, et al.
Published: (2025)
Intent-Driven Dynamic Chunking: Segmenting Documents to Reflect Predicted Information Needs
by: Koutsiaris, Christos
Published: (2026)
by: Koutsiaris, Christos
Published: (2026)
Exploring Information Retrieval Landscapes: An Investigation of a Novel Evaluation Techniques and Comparative Document Splitting Methods
by: Narimissa, Esmaeil, et al.
Published: (2024)
by: Narimissa, Esmaeil, et al.
Published: (2024)
MasterSet: A Large-Scale Benchmark for Must-Cite Citation Recommendation in the AI/ML Literature
by: Ratul, Md Toyaha Rahman, et al.
Published: (2026)
by: Ratul, Md Toyaha Rahman, et al.
Published: (2026)
BridgeRAG: Training-Free Bridge-Conditioned Retrieval for Multi-Hop Question Answering
by: Bacellar, Andre
Published: (2026)
by: Bacellar, Andre
Published: (2026)
Train Once, Use Flexibly: A Modular Framework for Multi-Aspect Neural News Recommendation
by: Iana, Andreea, et al.
Published: (2023)
by: Iana, Andreea, et al.
Published: (2023)
Session Context Embedding for Intent Understanding in Product Search
by: Mehrdad, Navid, et al.
Published: (2024)
by: Mehrdad, Navid, et al.
Published: (2024)
Scaling Multilingual Semantic Search in Uber Eats Delivery
by: Ling, Bo, et al.
Published: (2026)
by: Ling, Bo, et al.
Published: (2026)
Does UMBRELA Work on Other LLMs?
by: Farzi, Naghmeh, et al.
Published: (2025)
by: Farzi, Naghmeh, et al.
Published: (2025)
Algorithmic Trust and Compliance: Benchmarking Brand Notability for UK iGaming Entities in Generative Search Engines
by: Oruesagasti, Julen
Published: (2026)
by: Oruesagasti, Julen
Published: (2026)
Graph-GRPO: Dependency-Aware Credit Assignment for Generative E-commerce Search Relevance
by: Che, Jiarui, et al.
Published: (2026)
by: Che, Jiarui, et al.
Published: (2026)
Optimizing open-domain question answering with graph-based retrieval augmented generation
by: Cahoon, Joyce, et al.
Published: (2025)
by: Cahoon, Joyce, et al.
Published: (2025)
From Citation Selection to Citation Absorption: A Measurement Framework for Generative Engine Optimization Across AI Search Platforms
by: Kai, Zhang, et al.
Published: (2026)
by: Kai, Zhang, et al.
Published: (2026)
Interactive Question Answering Systems: Literature Review
by: Biancofiore, Giovanni Maria, et al.
Published: (2022)
by: Biancofiore, Giovanni Maria, et al.
Published: (2022)
Expanding Relevance Judgments for Medical Case-based Retrieval Task with Multimodal LLMs
by: Pires, Catarina, et al.
Published: (2025)
by: Pires, Catarina, et al.
Published: (2025)
What Matters in LLM-Based Feature Extractor for Recommender? A Systematic Analysis of Prompts, Models, and Adaptation
by: Shi, Kainan, et al.
Published: (2025)
by: Shi, Kainan, et al.
Published: (2025)
AuthorityBench: Benchmarking LLM Authority Perception for Reliable Retrieval-Augmented Generation
by: Yao, Zhihui, et al.
Published: (2026)
by: Yao, Zhihui, et al.
Published: (2026)
Less LLM, More Documents: Searching for Improved RAG
by: Ning, Jingjie, et al.
Published: (2025)
by: Ning, Jingjie, et al.
Published: (2025)
A Systematic Framework for Enterprise Knowledge Retrieval: Leveraging LLM-Generated Metadata to Enhance RAG Systems
by: Mishra, Pranav Pushkar, et al.
Published: (2025)
by: Mishra, Pranav Pushkar, et al.
Published: (2025)
Diversification as Risk Minimization
by: Takehi, Rikiya, et al.
Published: (2025)
by: Takehi, Rikiya, et al.
Published: (2025)
Evaluating the Effectiveness of Large Language Models in Automated News Article Summarization
by: Houamegni, Lionel Richy Panlap, et al.
Published: (2025)
by: Houamegni, Lionel Richy Panlap, et al.
Published: (2025)
Peeling Back the Layers: An In-Depth Evaluation of Encoder Architectures in Neural News Recommenders
by: Iana, Andreea, et al.
Published: (2024)
by: Iana, Andreea, et al.
Published: (2024)
A Case Study of Balanced Query Recommendation on Wikipedia
by: Mishra, Harshit, et al.
Published: (2025)
by: Mishra, Harshit, et al.
Published: (2025)
Harnessing multiple LLMs for Information Retrieval: A case study on Deep Learning methodologies in Biodiversity publications
by: Kommineni, Vamsi Krishna, et al.
Published: (2024)
by: Kommineni, Vamsi Krishna, et al.
Published: (2024)
Temporal Leakage in Search-Engine Date-Filtered Web Retrieval: A Retrospective Forecasting Case Study
by: Lahib, Ali El, et al.
Published: (2026)
by: Lahib, Ali El, et al.
Published: (2026)
Recall Isn't Enough: Bounding Commitments in Personalized Language Systems
by: Tang, Rui, et al.
Published: (2026)
by: Tang, Rui, et al.
Published: (2026)
Likert or Not: LLM Absolute Relevance Judgments on Fine-Grained Ordinal Scales
by: Godfrey, Charles, et al.
Published: (2025)
by: Godfrey, Charles, et al.
Published: (2025)
ECHO: An Open Research Platform for Evaluation of Chat, Human Behavior, and Outcomes
by: Liu, Jiqun, et al.
Published: (2026)
by: Liu, Jiqun, et al.
Published: (2026)
RobustExplain: Evaluating Robustness of LLM-Based Explanation Agents for Recommendation
by: Zhang, Guilin, et al.
Published: (2026)
by: Zhang, Guilin, et al.
Published: (2026)
VOGUE: A Multimodal Dataset for Conversational Recommendation in Fashion
by: Guo, David, et al.
Published: (2025)
by: Guo, David, et al.
Published: (2025)
LLM Reasoning for Cold-Start Item Recommendation
by: Li, Shijun, et al.
Published: (2025)
by: Li, Shijun, et al.
Published: (2025)
Augmented Relevance Datasets with Fine-Tuned Small LLMs
by: Fitte-Rey, Quentin, et al.
Published: (2025)
by: Fitte-Rey, Quentin, et al.
Published: (2025)
MLDocRAG: Multimodal Long-Context Document Retrieval Augmented Generation
by: Zhang, Yongyue, et al.
Published: (2026)
by: Zhang, Yongyue, et al.
Published: (2026)
Similar Items
-
Evaluating User Experience in Conversational Recommender Systems: A Systematic Review Across Classical and LLM-Powered Approaches
by: Mahmud, Raj, et al.
Published: (2025) -
Structural Feature Engineering for Generative Engine Optimization: How Content Structure Shapes Citation Behavior
by: Yu, Junwei, et al.
Published: (2026) -
From latent factors to language: a user study on LLM-generated explanations for an inherently interpretable matrix-based recommender system
by: Manderlier, Maxime, et al.
Published: (2025) -
Architecture Matters More Than Scale: A Comparative Study of Retrieval and Memory Augmentation for Financial QA Under SME Compute Constraints
by: Liu, Jianan, et al.
Published: (2026) -
Irec: A Metacognitive Scaffolding for Self-Regulated Learning through Just-in-Time Insight Recall: A Conceptual Framework and System Prototype
by: Hou, Xuefei, et al.
Published: (2025)