Trust, Safety, and Accuracy: Assessing LLMs for Routine Maternity Advice
Fuente:
arXiv
Guardado en:
| Autores principales: | Divya, V Sai, Bhanusree, A, Rimjhim, Rao, K Venkata Krishna |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Comparative Analysis of Large Language Models in Generating Telugu Responses for Maternal Health Queries
por: Bhanusree, Anagani, et al.
Publicado: (2026)
por: Bhanusree, Anagani, et al.
Publicado: (2026)
Recognition Without Authorization: LLMs and the Moral Order of Online Advice
por: van Nuenen, Tom
Publicado: (2026)
por: van Nuenen, Tom
Publicado: (2026)
SafeMath: Inference-time Safety improves Math Accuracy
por: Basu, Sagnik, et al.
Publicado: (2026)
por: Basu, Sagnik, et al.
Publicado: (2026)
Building Trust: Foundations of Security, Safety and Transparency in AI
por: Sidhpurwala, Huzaifa, et al.
Publicado: (2024)
por: Sidhpurwala, Huzaifa, et al.
Publicado: (2024)
Beyond Accuracy: Diagnosing Algebraic Reasoning Failures in LLMs Across Nine Complexity Dimensions
por: Patil, Parth, et al.
Publicado: (2026)
por: Patil, Parth, et al.
Publicado: (2026)
Ensuring Safety and Trust: Analyzing the Risks of Large Language Models in Medicine
por: Yang, Yifan, et al.
Publicado: (2024)
por: Yang, Yifan, et al.
Publicado: (2024)
Assessing the Impact of Conspiracy Theories Using Large Language Models
por: Jiang, Bohan, et al.
Publicado: (2024)
por: Jiang, Bohan, et al.
Publicado: (2024)
Leveraging LLMs to Assess Tutor Moves in Real-Life Dialogues: A Feasibility Study
por: Thomas, Danielle R., et al.
Publicado: (2025)
por: Thomas, Danielle R., et al.
Publicado: (2025)
Assessing the Performance of Human-Capable LLMs -- Are LLMs Coming for Your Job?
por: Mavi, John, et al.
Publicado: (2024)
por: Mavi, John, et al.
Publicado: (2024)
Help! Need Advice on Identifying Advice
por: Govindarajan, Venkata Subrahmanyan, et al.
Publicado: (2020)
por: Govindarajan, Venkata Subrahmanyan, et al.
Publicado: (2020)
Persuasion Dynamics in LLMs: Investigating Robustness and Adaptability in Knowledge and Safety with DuET-PD
por: Tan, Bryan Chen Zhengyu, et al.
Publicado: (2025)
por: Tan, Bryan Chen Zhengyu, et al.
Publicado: (2025)
The Reliability of LLMs for Medical Diagnosis: An Examination of Consistency, Manipulation, and Contextual Awareness
por: Subedi, Krishna
Publicado: (2025)
por: Subedi, Krishna
Publicado: (2025)
Exploring Safety Alignment Evaluation of LLMs in Chinese Mental Health Dialogues via LLM-as-Judge
por: Cai, Yunna, et al.
Publicado: (2025)
por: Cai, Yunna, et al.
Publicado: (2025)
When AI Speaks, Whose Values Does It Express? A Cross-Cultural Audit of Individualism-Collectivism Bias in Large Language Models
por: Venkata, Pruthvinath Jeripity
Publicado: (2026)
por: Venkata, Pruthvinath Jeripity
Publicado: (2026)
Are LLMs Court-Ready? Evaluating Frontier Models on Indian Legal Reasoning
por: Juvekar, Kush, et al.
Publicado: (2025)
por: Juvekar, Kush, et al.
Publicado: (2025)
Can LLMs Reason About Trust?: A Pilot Study
por: Debnath, Anushka, et al.
Publicado: (2025)
por: Debnath, Anushka, et al.
Publicado: (2025)
Assessing LLMs in Art Contexts: Critique Generation and Theory of Mind Evaluation
por: Arita, Takaya, et al.
Publicado: (2025)
por: Arita, Takaya, et al.
Publicado: (2025)
The Homogenization Problem in LLMs: Towards Meaningful Diversity in AI Safety
por: Rios-Sialer, Ian
Publicado: (2026)
por: Rios-Sialer, Ian
Publicado: (2026)
When Can We Trust LLMs in Mental Health? Large-Scale Benchmarks for Reliable LLM Evaluation
por: Badawi, Abeer, et al.
Publicado: (2025)
por: Badawi, Abeer, et al.
Publicado: (2025)
ZPD-SCA: Unveiling the Blind Spots of LLMs in Assessing Students' Cognitive Abilities
por: Dong, Wenhan, et al.
Publicado: (2025)
por: Dong, Wenhan, et al.
Publicado: (2025)
Real or Robotic? Assessing Whether LLMs Accurately Simulate Qualities of Human Responses in Dialogue
por: Ivey, Jonathan, et al.
Publicado: (2024)
por: Ivey, Jonathan, et al.
Publicado: (2024)
LLM or Human? Perceptions of Trust and Information Quality in Research Summaries
por: Akpinar, Nil-Jana, et al.
Publicado: (2026)
por: Akpinar, Nil-Jana, et al.
Publicado: (2026)
Counterfactual Probing for the Influence of Affect and Specificity on Intergroup Bias
por: Govindarajan, Venkata S, et al.
Publicado: (2023)
por: Govindarajan, Venkata S, et al.
Publicado: (2023)
SafetyAnalyst: Interpretable, Transparent, and Steerable Safety Moderation for AI Behavior
por: Li, Jing-Jing, et al.
Publicado: (2024)
por: Li, Jing-Jing, et al.
Publicado: (2024)
Evaluating Prompt Engineering Techniques for Accuracy and Confidence Elicitation in Medical LLMs
por: Naderi, Nariman, et al.
Publicado: (2025)
por: Naderi, Nariman, et al.
Publicado: (2025)
Passing the Turing Test in Political Discourse: Fine-Tuning LLMs to Mimic Polarized Social Media Comments
por: Pazzaglia, ., et al.
Publicado: (2025)
por: Pazzaglia, ., et al.
Publicado: (2025)
CAIRNS: Balancing Readability and Scientific Accuracy in Climate Adaptation Question Answering
por: Kong, Liangji, et al.
Publicado: (2025)
por: Kong, Liangji, et al.
Publicado: (2025)
The Statistical Signature of LLMs
por: Hadad, Ortal, et al.
Publicado: (2026)
por: Hadad, Ortal, et al.
Publicado: (2026)
Different Bias Under Different Criteria: Assessing Bias in LLMs with a Fact-Based Approach
por: Ko, Changgeon, et al.
Publicado: (2024)
por: Ko, Changgeon, et al.
Publicado: (2024)
When Do LLMs Generate Realistic Social Networks? A Multi-Dimensional Study of Culture, Language, Scale, and Method
por: Kilaru, Sai Hemanth, et al.
Publicado: (2026)
por: Kilaru, Sai Hemanth, et al.
Publicado: (2026)
When Can We Trust LLM Graders? Calibrating Confidence for Automated Assessment
por: Ferrer, Robinson, et al.
Publicado: (2026)
por: Ferrer, Robinson, et al.
Publicado: (2026)
Words of Warmth: Trust and Sociability Norms for over 26k English Words
por: Mohammad, Saif M.
Publicado: (2025)
por: Mohammad, Saif M.
Publicado: (2025)
Auditing Agent Harness Safety
por: Liu, Chengzhi, et al.
Publicado: (2026)
por: Liu, Chengzhi, et al.
Publicado: (2026)
LLMs left, right, and center: Assessing GPT's capabilities to label political bias from web domains
por: Hernandes, Raphael, et al.
Publicado: (2024)
por: Hernandes, Raphael, et al.
Publicado: (2024)
Paradox of De-identification: A Critique of HIPAA Safe Harbour in the Age of LLMs
por: Jiang, Lavender Y., et al.
Publicado: (2026)
por: Jiang, Lavender Y., et al.
Publicado: (2026)
Unfair TOS: An Automated Approach using Customized BERT
por: Akash, Bathini Sai, et al.
Publicado: (2024)
por: Akash, Bathini Sai, et al.
Publicado: (2024)
Expected Harm: Rethinking Safety Evaluation of (Mis)Aligned LLMs
por: Chen, Yen-Shan, et al.
Publicado: (2026)
por: Chen, Yen-Shan, et al.
Publicado: (2026)
Assessing Judging Bias in Large Reasoning Models: An Empirical Study
por: Wang, Qian, et al.
Publicado: (2025)
por: Wang, Qian, et al.
Publicado: (2025)
Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset
por: Chehbouni, Khaoula, et al.
Publicado: (2024)
por: Chehbouni, Khaoula, et al.
Publicado: (2024)
Hallucination-Free? Assessing the Reliability of Leading AI Legal Research Tools
por: Magesh, Varun, et al.
Publicado: (2024)
por: Magesh, Varun, et al.
Publicado: (2024)
Ejemplares similares
-
Comparative Analysis of Large Language Models in Generating Telugu Responses for Maternal Health Queries
por: Bhanusree, Anagani, et al.
Publicado: (2026) -
Recognition Without Authorization: LLMs and the Moral Order of Online Advice
por: van Nuenen, Tom
Publicado: (2026) -
SafeMath: Inference-time Safety improves Math Accuracy
por: Basu, Sagnik, et al.
Publicado: (2026) -
Building Trust: Foundations of Security, Safety and Transparency in AI
por: Sidhpurwala, Huzaifa, et al.
Publicado: (2024) -
Beyond Accuracy: Diagnosing Algebraic Reasoning Failures in LLMs Across Nine Complexity Dimensions
por: Patil, Parth, et al.
Publicado: (2026)