How effective are VLMs in assisting humans in inferring the quality of mental models from Multimodal short answers?
Fuente:
arXiv
Saved in:
| Main Authors: | Sil, Pritam, Karnam, Durgaprasad, Venumuddala, Vinay Reddy, Bhattacharyya, Pushpak |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
"I understand why I got this grade": Automatic Short Answer Grading with Feedback
by: Aggarwal, Dishank, et al.
Published: (2024)
by: Aggarwal, Dishank, et al.
Published: (2024)
Can MLLMs generate human-like feedback in grading multimodal short answers?
by: Sil, Pritam, et al.
Published: (2024)
by: Sil, Pritam, et al.
Published: (2024)
StereoDetect: Detecting Stereotypes and Anti-stereotypes the Correct Way Using Social Psychological Underpinnings
by: Shejole, Kaustubh Shivshankar, et al.
Published: (2025)
by: Shejole, Kaustubh Shivshankar, et al.
Published: (2025)
Machine-assisted quantitizing designs: augmenting humanities and social sciences with artificial intelligence
by: Karjus, Andres
Published: (2023)
by: Karjus, Andres
Published: (2023)
Rethinking Test-Time Scaling for Medical AI: Model and Task-Aware Strategies for LLMs and VLMs
by: Oh, Gyutaek, et al.
Published: (2025)
by: Oh, Gyutaek, et al.
Published: (2025)
ToxVidLM: A Multimodal Framework for Toxicity Detection in Code-Mixed Videos
by: Maity, Krishanu, et al.
Published: (2024)
by: Maity, Krishanu, et al.
Published: (2024)
A closer look at how large language models trust humans: patterns and biases
by: Lerman, Valeria, et al.
Published: (2025)
by: Lerman, Valeria, et al.
Published: (2025)
EthicsMH: A Pilot Benchmark for Ethical Reasoning in Mental Health AI
by: Kasu, Sai Kartheek Reddy
Published: (2025)
by: Kasu, Sai Kartheek Reddy
Published: (2025)
Recon, Answer, Verify: Agents in Search of Truth
by: Shukla, Satyam, et al.
Published: (2025)
by: Shukla, Satyam, et al.
Published: (2025)
We Care: Multimodal Depression Detection and Knowledge Infused Mental Health Therapeutic Response Generation
by: Moon, Palash, et al.
Published: (2024)
by: Moon, Palash, et al.
Published: (2024)
CEGI: Measuring the trade-off between efficiency and carbon emissions for SLMs and VLMs
by: Kumar, Abhas, et al.
Published: (2024)
by: Kumar, Abhas, et al.
Published: (2024)
The Sensation Modulating Network:Haltability as the architectural ground for object-directed phenomenology
by: Nagarjuna, G., et al.
Published: (2026)
by: Nagarjuna, G., et al.
Published: (2026)
Enhancing reasoning accuracy in large language models during inference time
by: Sharma, Vinay, et al.
Published: (2026)
by: Sharma, Vinay, et al.
Published: (2026)
Looks can be Deceptive: Distinguishing Repetition Disfluency from Reduplication
by: Ahmad, Arif, et al.
Published: (2024)
by: Ahmad, Arif, et al.
Published: (2024)
Large language models can consistently generate high-quality content for election disinformation operations
by: Williams, Angus R., et al.
Published: (2024)
by: Williams, Angus R., et al.
Published: (2024)
Does GPT-4 surpass human performance in linguistic pragmatics?
by: Bojic, Ljubisa, et al.
Published: (2023)
by: Bojic, Ljubisa, et al.
Published: (2023)
HumT DumT: Measuring and controlling human-like language in LLMs
by: Cheng, Myra, et al.
Published: (2025)
by: Cheng, Myra, et al.
Published: (2025)
How effective is Multi-source pivoting for Translation of Low Resource Indian Languages?
by: Gaikwad, Pranav, et al.
Published: (2024)
by: Gaikwad, Pranav, et al.
Published: (2024)
How Large Language Models play humans in online conversations: a simulated study of the 2016 US politics on Reddit
by: Cirulli, Daniele, et al.
Published: (2025)
by: Cirulli, Daniele, et al.
Published: (2025)
Enhancing Food-Domain Question Answering with a Multimodal Knowledge Graph: Hybrid QA Generation and Diversity Analysis
by: B, Srihari K, et al.
Published: (2025)
by: B, Srihari K, et al.
Published: (2025)
Taking a turn for the better: Conversation redirection throughout the course of mental-health therapy
by: Nguyen, Vivian, et al.
Published: (2024)
by: Nguyen, Vivian, et al.
Published: (2024)
Unveiling the Invisible: Captioning Videos with Metaphors
by: Kalarani, Abisek Rajakumar, et al.
Published: (2024)
by: Kalarani, Abisek Rajakumar, et al.
Published: (2024)
How Large Language Models are Designed to Hallucinate
by: Ackermann, Richard, et al.
Published: (2025)
by: Ackermann, Richard, et al.
Published: (2025)
The opportunities and risks of large language models in mental health
by: Lawrence, Hannah R., et al.
Published: (2024)
by: Lawrence, Hannah R., et al.
Published: (2024)
EduIllustrate: Towards Scalable Automated Generation Of Multimodal Educational Content
by: Bi, Shuzhen, et al.
Published: (2026)
by: Bi, Shuzhen, et al.
Published: (2026)
Exploring Multimodal Challenges in Toxic Chinese Detection: Taxonomy, Benchmark, and Findings
by: Yang, Shujian, et al.
Published: (2025)
by: Yang, Shujian, et al.
Published: (2025)
AI Act and Large Language Models (LLMs): When critical issues and privacy impact require human and ethical oversight
by: Fabiano, Nicola
Published: (2024)
by: Fabiano, Nicola
Published: (2024)
The Last Fingerprint: How Markdown Training Shapes LLM Prose
by: Freeburg, E. M.
Published: (2026)
by: Freeburg, E. M.
Published: (2026)
The Incomplete Bridge: How AI Research (Mis)Engages with Psychology
by: Jiang, Han, et al.
Published: (2025)
by: Jiang, Han, et al.
Published: (2025)
How Did We Get Here? Summarizing Conversation Dynamics
by: Hua, Yilun, et al.
Published: (2024)
by: Hua, Yilun, et al.
Published: (2024)
MATCHED: Multimodal Authorship-Attribution To Combat Human Trafficking in Escort-Advertisement Data
by: Saxena, Vageesh, et al.
Published: (2024)
by: Saxena, Vageesh, et al.
Published: (2024)
Asking an AI for salary negotiation advice is a matter of concern: Controlled experimental perturbation of ChatGPT for protected and non-protected group discrimination on a contextual task with no clear ground truth answers
by: Geiger, R. Stuart, et al.
Published: (2024)
by: Geiger, R. Stuart, et al.
Published: (2024)
How English Print Media Frames Human-Elephant Conflicts in India
by: Punith, Bonala Sai, et al.
Published: (2026)
by: Punith, Bonala Sai, et al.
Published: (2026)
From Text to Multimodality: Exploring the Evolution and Impact of Large Language Models in Medical Practice
by: Niu, Qian, et al.
Published: (2024)
by: Niu, Qian, et al.
Published: (2024)
Cognitive Chain-of-Thought (CoCoT): Structured Multimodal Reasoning about Social Situations
by: Park, Eunkyu, et al.
Published: (2025)
by: Park, Eunkyu, et al.
Published: (2025)
All in How You Ask for It: Simple Black-Box Method for Jailbreak Attacks
by: Takemoto, Kazuhiro
Published: (2024)
by: Takemoto, Kazuhiro
Published: (2024)
Autoscoring Anticlimax: A Meta-analytic Understanding of AI's Short-answer Shortcomings and Wording Weaknesses
by: Hardy, Michael
Published: (2026)
by: Hardy, Michael
Published: (2026)
"Sorry, I Didn't Catch That": How Speech Models Miss What Matters Most
by: Zhou, Kaitlyn, et al.
Published: (2026)
by: Zhou, Kaitlyn, et al.
Published: (2026)
The Silent Curriculum: How Does LLM Monoculture Shape Educational Content and Its Accessibility?
by: Priyanshu, Aman, et al.
Published: (2024)
by: Priyanshu, Aman, et al.
Published: (2024)
Ads in AI Chatbots? An Analysis of How Large Language Models Navigate Conflicts of Interest
by: Wu, Addison J., et al.
Published: (2026)
by: Wu, Addison J., et al.
Published: (2026)
Similar Items
-
"I understand why I got this grade": Automatic Short Answer Grading with Feedback
by: Aggarwal, Dishank, et al.
Published: (2024) -
Can MLLMs generate human-like feedback in grading multimodal short answers?
by: Sil, Pritam, et al.
Published: (2024) -
StereoDetect: Detecting Stereotypes and Anti-stereotypes the Correct Way Using Social Psychological Underpinnings
by: Shejole, Kaustubh Shivshankar, et al.
Published: (2025) -
Machine-assisted quantitizing designs: augmenting humanities and social sciences with artificial intelligence
by: Karjus, Andres
Published: (2023) -
Rethinking Test-Time Scaling for Medical AI: Model and Task-Aware Strategies for LLMs and VLMs
by: Oh, Gyutaek, et al.
Published: (2025)