Accuracy and Consistency of LLMs in the Registered Dietitian Exam: The Impact of Prompt Engineering and Knowledge Retrieval
Fuente:
arXiv
Salvato in:
| Autori principali: | Azimi, Iman, Qi, Mohan, Wang, Li, Rahmani, Amir M., Li, Youlin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
An LLM-Powered Agent for Physiological Data Analysis: A Case Study on PPG-based Heart Rate Estimation
di: Feli, Mohammad, et al.
Pubblicazione: (2025)
di: Feli, Mohammad, et al.
Pubblicazione: (2025)
Conversational Health Agents: A Personalized LLM-Powered Agent Framework
di: Abbasian, Mahyar, et al.
Pubblicazione: (2023)
di: Abbasian, Mahyar, et al.
Pubblicazione: (2023)
Empathy Through Multimodality in Conversational Interfaces
di: Abbasian, Mahyar, et al.
Pubblicazione: (2024)
di: Abbasian, Mahyar, et al.
Pubblicazione: (2024)
Knowledge-Infused LLM-Powered Conversational Health Agent: A Case Study for Diabetes Patients
di: Abbasian, Mahyar, et al.
Pubblicazione: (2024)
di: Abbasian, Mahyar, et al.
Pubblicazione: (2024)
ALCM: Autonomous LLM-Augmented Causal Discovery Framework
di: Khatibi, Elahe, et al.
Pubblicazione: (2024)
di: Khatibi, Elahe, et al.
Pubblicazione: (2024)
Evaluating Prompt Engineering Techniques for Accuracy and Confidence Elicitation in Medical LLMs
di: Naderi, Nariman, et al.
Pubblicazione: (2025)
di: Naderi, Nariman, et al.
Pubblicazione: (2025)
Personalized Causal Graph Reasoning for LLMs: An Implementation for Dietary Recommendations
di: Yang, Zhongqi, et al.
Pubblicazione: (2025)
di: Yang, Zhongqi, et al.
Pubblicazione: (2025)
Lifelong Knowledge Editing for LLMs with Retrieval-Augmented Continuous Prompt Learning
di: Chen, Qizhou, et al.
Pubblicazione: (2024)
di: Chen, Qizhou, et al.
Pubblicazione: (2024)
CDF-RAG: Causal Dynamic Feedback for Adaptive Retrieval-Augmented Generation
di: Khatibi, Elahe, et al.
Pubblicazione: (2025)
di: Khatibi, Elahe, et al.
Pubblicazione: (2025)
Consistency Guided Knowledge Retrieval and Denoising in LLMs for Zero-shot Document-level Relation Triplet Extraction
di: Sun, Qi, et al.
Pubblicazione: (2024)
di: Sun, Qi, et al.
Pubblicazione: (2024)
TACOMORE: Leveraging the Potential of LLMs in Corpus-based Discourse Analysis with Prompt Engineering
di: Li, Bingru, et al.
Pubblicazione: (2024)
di: Li, Bingru, et al.
Pubblicazione: (2024)
Graph-Augmented LLMs for Personalized Health Insights: A Case Study in Sleep Analysis
di: Subramanian, Ajan, et al.
Pubblicazione: (2024)
di: Subramanian, Ajan, et al.
Pubblicazione: (2024)
Beyond Self-Consistency: Ensemble Reasoning Boosts Consistency and Accuracy of LLMs in Cancer Staging
di: Chang, Chia-Hsuan, et al.
Pubblicazione: (2024)
di: Chang, Chia-Hsuan, et al.
Pubblicazione: (2024)
Emulating Retrieval Augmented Generation via Prompt Engineering for Enhanced Long Context Comprehension in LLMs
di: Park, Joon, et al.
Pubblicazione: (2025)
di: Park, Joon, et al.
Pubblicazione: (2025)
Foundation Metrics for Evaluating Effectiveness of Healthcare Conversations Powered by Generative AI
di: Abbasian, Mahyar, et al.
Pubblicazione: (2023)
di: Abbasian, Mahyar, et al.
Pubblicazione: (2023)
Confidence Improves Self-Consistency in LLMs
di: Taubenfeld, Amir, et al.
Pubblicazione: (2025)
di: Taubenfeld, Amir, et al.
Pubblicazione: (2025)
Prompt Engineering: How Prompt Vocabulary affects Domain Knowledge
di: Schreiter, Dimitri
Pubblicazione: (2025)
di: Schreiter, Dimitri
Pubblicazione: (2025)
Reverse Prompt Engineering
di: Li, Hanqing, et al.
Pubblicazione: (2024)
di: Li, Hanqing, et al.
Pubblicazione: (2024)
The Ever-Evolving Science Exam
di: Wang, Junying, et al.
Pubblicazione: (2025)
di: Wang, Junying, et al.
Pubblicazione: (2025)
Dissecting Paraphrases: The Impact of Prompt Syntax and supplementary Information on Knowledge Retrieval from Pretrained Language Models
di: Linzbach, Stephan, et al.
Pubblicazione: (2024)
di: Linzbach, Stephan, et al.
Pubblicazione: (2024)
MedREK: Retrieval-Based Editing for Medical LLMs with Key-Aware Prompts
di: Xia, Shujun, et al.
Pubblicazione: (2025)
di: Xia, Shujun, et al.
Pubblicazione: (2025)
How Well Do LLMs Understand Drug Mechanisms? A Knowledge + Reasoning Evaluation Dataset
di: Mohan, Sunil, et al.
Pubblicazione: (2025)
di: Mohan, Sunil, et al.
Pubblicazione: (2025)
MedCoT-RAG: Causal Chain-of-Thought RAG for Medical Question Answering
di: Wang, Ziyu, et al.
Pubblicazione: (2025)
di: Wang, Ziyu, et al.
Pubblicazione: (2025)
Do LLMs have Consistent Values?
di: Rozen, Naama, et al.
Pubblicazione: (2024)
di: Rozen, Naama, et al.
Pubblicazione: (2024)
Can We Afford The Perfect Prompt? Balancing Cost and Accuracy with the Economical Prompting Index
di: McDonald, Tyler, et al.
Pubblicazione: (2024)
di: McDonald, Tyler, et al.
Pubblicazione: (2024)
A Retrieval-Augmented Knowledge Mining Method with Deep Thinking LLMs for Biomedical Research and Clinical Support
di: Feng, Yichun, et al.
Pubblicazione: (2025)
di: Feng, Yichun, et al.
Pubblicazione: (2025)
Aligning What LLMs Do and Say: Towards Self-Consistent Explanations
di: Admoni, Sahar, et al.
Pubblicazione: (2025)
di: Admoni, Sahar, et al.
Pubblicazione: (2025)
Enhancing Knowledge Distillation for LLMs with Response-Priming Prompting
di: Goyal, Vijay, et al.
Pubblicazione: (2024)
di: Goyal, Vijay, et al.
Pubblicazione: (2024)
MegaChat: A Synthetic Persian Q&A Dataset for High-Quality Sales Chatbot Evaluation
di: Rahmani, Mahdi, et al.
Pubblicazione: (2025)
di: Rahmani, Mahdi, et al.
Pubblicazione: (2025)
Skewed Memorization in Large Language Models: Quantification and Decomposition
di: Li, Hao, et al.
Pubblicazione: (2025)
di: Li, Hao, et al.
Pubblicazione: (2025)
Last Layer Logits to Logic: Empowering LLMs with Logic-Consistent Structured Knowledge Reasoning
di: Li, Songze, et al.
Pubblicazione: (2025)
di: Li, Songze, et al.
Pubblicazione: (2025)
Too Consistent to Detect: A Study of Self-Consistent Errors in LLMs
di: Tan, Hexiang, et al.
Pubblicazione: (2025)
di: Tan, Hexiang, et al.
Pubblicazione: (2025)
Integration of Old and New Knowledge for Generalized Intent Discovery: A Consistency-driven Prototype-Prompting Framework
di: Wei, Xiao, et al.
Pubblicazione: (2025)
di: Wei, Xiao, et al.
Pubblicazione: (2025)
Attention Consistency for LLMs Explanation
di: Lan, Tian, et al.
Pubblicazione: (2025)
di: Lan, Tian, et al.
Pubblicazione: (2025)
AVMeme Exam: A Multimodal Multilingual Multicultural Benchmark for LLMs' Contextual and Cultural Knowledge and Thinking
di: Jiang, Xilin, et al.
Pubblicazione: (2026)
di: Jiang, Xilin, et al.
Pubblicazione: (2026)
Transforming External Knowledge into Triplets for Enhanced Retrieval in RAG of LLMs
di: Wang, Xudong, et al.
Pubblicazione: (2026)
di: Wang, Xudong, et al.
Pubblicazione: (2026)
Humanity's Last Code Exam: Can Advanced LLMs Conquer Human's Hardest Code Competition?
di: Li, Xiangyang, et al.
Pubblicazione: (2025)
di: Li, Xiangyang, et al.
Pubblicazione: (2025)
Accommodate Knowledge Conflicts in Retrieval-augmented LLMs: Towards Robust Response Generation in the Wild
di: Wang, Jiatai, et al.
Pubblicazione: (2025)
di: Wang, Jiatai, et al.
Pubblicazione: (2025)
LongIns: A Challenging Long-context Instruction-based Exam for LLMs
di: Gavin, Shawn, et al.
Pubblicazione: (2024)
di: Gavin, Shawn, et al.
Pubblicazione: (2024)
Differential Private Federated Transfer Learning for Mental Health Monitoring in Everyday Settings: A Case Study on Stress Detection
di: Wang, Ziyu, et al.
Pubblicazione: (2024)
di: Wang, Ziyu, et al.
Pubblicazione: (2024)
Documenti analoghi
-
An LLM-Powered Agent for Physiological Data Analysis: A Case Study on PPG-based Heart Rate Estimation
di: Feli, Mohammad, et al.
Pubblicazione: (2025) -
Conversational Health Agents: A Personalized LLM-Powered Agent Framework
di: Abbasian, Mahyar, et al.
Pubblicazione: (2023) -
Empathy Through Multimodality in Conversational Interfaces
di: Abbasian, Mahyar, et al.
Pubblicazione: (2024) -
Knowledge-Infused LLM-Powered Conversational Health Agent: A Case Study for Diabetes Patients
di: Abbasian, Mahyar, et al.
Pubblicazione: (2024) -
ALCM: Autonomous LLM-Augmented Causal Discovery Framework
di: Khatibi, Elahe, et al.
Pubblicazione: (2024)