CLINIC: Evaluating Multilingual Trustworthiness in Language Models for Healthcare
Fuente:
arXiv
Saved in:
| Main Authors: | Ghosh, Akash, Sridhar, Srivarshinee, Ravi, Raghav Kaushik, Muhsin, Muhsin, Saha, Sriparna, Agarwal, Chirag |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Survey of Multilingual Reasoning in Language Models
by: Ghosh, Akash, et al.
Published: (2025)
by: Ghosh, Akash, et al.
Published: (2025)
Mapping Clinical Doubt: Locating Linguistic Uncertainty in LLMs
by: Sridhar, Srivarshinee, et al.
Published: (2025)
by: Sridhar, Srivarshinee, et al.
Published: (2025)
CURE-Med: Curriculum-Informed Reinforcement Learning for Multilingual Medical Reasoning
by: Onyame, Eric, et al.
Published: (2026)
by: Onyame, Eric, et al.
Published: (2026)
SANSKRITI: A Comprehensive Benchmark for Evaluating Language Models' Knowledge of Indian Culture
by: Maji, Arijit, et al.
Published: (2025)
by: Maji, Arijit, et al.
Published: (2025)
DRISHTIKON: A Multimodal Multilingual Benchmark for Testing Language Models' Understanding on Indian Culture
by: Maji, Arijit, et al.
Published: (2025)
by: Maji, Arijit, et al.
Published: (2025)
Exploring the Frontier of Vision-Language Models: A Survey of Current Methodologies and Future Directions
by: Ghosh, Akash, et al.
Published: (2024)
by: Ghosh, Akash, et al.
Published: (2024)
Infogen: Generating Complex Statistical Infographics from Documents
by: Ghosh, Akash, et al.
Published: (2025)
by: Ghosh, Akash, et al.
Published: (2025)
MedSumm: A Multimodal Approach to Summarizing Code-Mixed Hindi-English Clinical Queries
by: Ghosh, Akash, et al.
Published: (2024)
by: Ghosh, Akash, et al.
Published: (2024)
A Comprehensive Survey of Hallucination in Large Language, Image, Video and Audio Foundation Models
by: Sahoo, Pranab, et al.
Published: (2024)
by: Sahoo, Pranab, et al.
Published: (2024)
Relic: Enhancing Reward Model Generalization for Low-Resource Indic Languages with Few-Shot Examples
by: Ghosal, Soumya Suvra, et al.
Published: (2025)
by: Ghosal, Soumya Suvra, et al.
Published: (2025)
Explain Thyself Bully: Sentiment Aided Cyberbullying Detection with Explanation
by: Maity, Krishanu, et al.
Published: (2024)
by: Maity, Krishanu, et al.
Published: (2024)
Hi Model, generating 'nice' instead of 'good' is not as bad as generating 'rice'! Towards Context and Semantic Infused Dialogue Generation Loss Function and Evaluation Metric
by: Tiwari, Abhisek, et al.
Published: (2023)
by: Tiwari, Abhisek, et al.
Published: (2023)
Assessing Empathy in Large Language Models with Real-World Physician-Patient Interactions
by: Luo, Man, et al.
Published: (2024)
by: Luo, Man, et al.
Published: (2024)
XTRUST: On the Multilingual Trustworthiness of Large Language Models
by: Li, Yahan, et al.
Published: (2024)
by: Li, Yahan, et al.
Published: (2024)
From Generation to Collaboration: Using LLMs to Edit for Empathy in Healthcare
by: Luo, Man, et al.
Published: (2026)
by: Luo, Man, et al.
Published: (2026)
Let's Play Across Cultures: A Large Multilingual, Multicultural Benchmark for Assessing Language Models' Understanding of Sports
by: Singh, Punit Kumar, et al.
Published: (2025)
by: Singh, Punit Kumar, et al.
Published: (2025)
Meme-ingful Analysis: Enhanced Understanding of Cyberbullying in Memes Through Multimodal Explanations
by: Jha, Prince, et al.
Published: (2024)
by: Jha, Prince, et al.
Published: (2024)
MemeGuard: An LLM and VLM-based Framework for Advancing Content Moderation via Meme Intervention
by: Jha, Prince, et al.
Published: (2024)
by: Jha, Prince, et al.
Published: (2024)
When Background Matters: Breaking Medical Vision Language Models by Transferable Attack
by: Ghosh, Akash, et al.
Published: (2026)
by: Ghosh, Akash, et al.
Published: (2026)
RFID-Based Non-Biometric Classroom Attendance System: Proxy Attendance Detection via Weight Sensor Integration
by: Ege, Furkan, et al.
Published: (2026)
by: Ege, Furkan, et al.
Published: (2026)
A Survey on Medical Document Summarization: From Machine Learning Techniques to Large Language Models
by: Akash Ghosh, et al.
Published: (2025)
by: Akash Ghosh, et al.
Published: (2025)
ReflectSumm: A Benchmark for Course Reflection Summarization
by: Zhong, Yang, et al.
Published: (2024)
by: Zhong, Yang, et al.
Published: (2024)
Polarity-Aware Probing for Quantifying Latent Alignment in Language Models
by: Sadiekh, Sabrina, et al.
Published: (2025)
by: Sadiekh, Sabrina, et al.
Published: (2025)
Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models
by: Agarwal, Chirag, et al.
Published: (2024)
by: Agarwal, Chirag, et al.
Published: (2024)
Two eyes, Two views, and finally, One summary! Towards Multi-modal Multi-tasking Knowledge-Infused Medical Dialogue Summarization
by: Saha, Anisha, et al.
Published: (2024)
by: Saha, Anisha, et al.
Published: (2024)
BhashaSutra: A Task-Centric Unified Survey of Indian NLP Datasets, Corpora, and Resources
by: Kumar, Raghvendra, et al.
Published: (2026)
by: Kumar, Raghvendra, et al.
Published: (2026)
SynopticBench: Evaluating Vision-Language Models on Generating Weather Forecast Discussions of the Future
by: Higgins, Timothy B., et al.
Published: (2026)
by: Higgins, Timothy B., et al.
Published: (2026)
A Comprehensive Survey on the Trustworthiness of Large Language Models in Healthcare
by: Aljohani, Manar, et al.
Published: (2025)
by: Aljohani, Manar, et al.
Published: (2025)
Yes, this is what I was looking for! Towards Multi-modal Medical Consultation Concern Summary Generation
by: Tiwari, Abhisek, et al.
Published: (2024)
by: Tiwari, Abhisek, et al.
Published: (2024)
VIRAASAT: Traversing Novel Paths for Indian Cultural Reasoning
by: Surana, Harshul Raj, et al.
Published: (2026)
by: Surana, Harshul Raj, et al.
Published: (2026)
On the Hardness of Faithful Chain-of-Thought Reasoning in Large Language Models
by: Tanneru, Sree Harsha, et al.
Published: (2024)
by: Tanneru, Sree Harsha, et al.
Published: (2024)
Towards Knowledge-Infused Automated Disease Diagnosis Assistant
by: Tomar, Mohit, et al.
Published: (2024)
by: Tomar, Mohit, et al.
Published: (2024)
COSMMIC: Comment-Sensitive Multimodal Multilingual Indian Corpus for Summarization and Headline Generation
by: Kumar, Raghvendra, et al.
Published: (2025)
by: Kumar, Raghvendra, et al.
Published: (2025)
Learning to Extract Cross-Domain Aspects and Understanding Sentiments Using Large Language Models
by: Ghosh, Karukriti Kaushik, et al.
Published: (2025)
by: Ghosh, Karukriti Kaushik, et al.
Published: (2025)
FIND: Toward Multimodal Financial Reasoning and Question Answering for Indic Languages
by: Das, Sarmistha, et al.
Published: (2026)
by: Das, Sarmistha, et al.
Published: (2026)
Krutrim LLM: Multilingual Foundational Model for over a Billion People
by: Kallappa, Aditya, et al.
Published: (2025)
by: Kallappa, Aditya, et al.
Published: (2025)
A Graph Talks, But Who's Listening? Rethinking Evaluations for Graph-Language Models
by: Petkar, Soham, et al.
Published: (2025)
by: Petkar, Soham, et al.
Published: (2025)
When Meaning Isn't Literal: Exploring Idiomatic Meaning Across Languages and Modalities
by: Das, Sarmistha, et al.
Published: (2026)
by: Das, Sarmistha, et al.
Published: (2026)
An EcoSage Assistant: Towards Building A Multimodal Plant Care Dialogue Assistant
by: Tomar, Mohit, et al.
Published: (2024)
by: Tomar, Mohit, et al.
Published: (2024)
Aligning Large Language Models with Healthcare Stakeholders: A Pathway to Trustworthy AI Integration
by: Ding, Kexin, et al.
Published: (2025)
by: Ding, Kexin, et al.
Published: (2025)
Similar Items
-
A Survey of Multilingual Reasoning in Language Models
by: Ghosh, Akash, et al.
Published: (2025) -
Mapping Clinical Doubt: Locating Linguistic Uncertainty in LLMs
by: Sridhar, Srivarshinee, et al.
Published: (2025) -
CURE-Med: Curriculum-Informed Reinforcement Learning for Multilingual Medical Reasoning
by: Onyame, Eric, et al.
Published: (2026) -
SANSKRITI: A Comprehensive Benchmark for Evaluating Language Models' Knowledge of Indian Culture
by: Maji, Arijit, et al.
Published: (2025) -
DRISHTIKON: A Multimodal Multilingual Benchmark for Testing Language Models' Understanding on Indian Culture
by: Maji, Arijit, et al.
Published: (2025)