"I understand why I got this grade": Automatic Short Answer Grading with Feedback
Fuente:
arXiv
Salvato in:
| Autori principali: | Aggarwal, Dishank, Sil, Pritam, Raman, Bhaskaran, Bhattacharyya, Pushpak |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
How effective are VLMs in assisting humans in inferring the quality of mental models from Multimodal short answers?
di: Sil, Pritam, et al.
Pubblicazione: (2026)
di: Sil, Pritam, et al.
Pubblicazione: (2026)
Can AI Assistance Aid in the Grading of Handwritten Answer Sheets?
di: Sil, Pritam, et al.
Pubblicazione: (2024)
di: Sil, Pritam, et al.
Pubblicazione: (2024)
Recon, Answer, Verify: Agents in Search of Truth
di: Shukla, Satyam, et al.
Pubblicazione: (2025)
di: Shukla, Satyam, et al.
Pubblicazione: (2025)
I've got the "Answer"! Interpretation of LLMs Hidden States in Question Answering
di: Goloviznina, Valeriya, et al.
Pubblicazione: (2024)
di: Goloviznina, Valeriya, et al.
Pubblicazione: (2024)
Can MLLMs generate human-like feedback in grading multimodal short answers?
di: Sil, Pritam, et al.
Pubblicazione: (2024)
di: Sil, Pritam, et al.
Pubblicazione: (2024)
Beyond Scores: A Modular RAG-Based System for Automatic Short Answer Scoring with Feedback
di: Fateen, Menna, et al.
Pubblicazione: (2024)
di: Fateen, Menna, et al.
Pubblicazione: (2024)
The why, what, and how of AI-based coding in scientific research
di: Zhuang, Tonghe, et al.
Pubblicazione: (2024)
di: Zhuang, Tonghe, et al.
Pubblicazione: (2024)
Yes, this is what I was looking for! Towards Multi-modal Medical Consultation Concern Summary Generation
di: Tiwari, Abhisek, et al.
Pubblicazione: (2024)
di: Tiwari, Abhisek, et al.
Pubblicazione: (2024)
ELEPHANT: Measuring and understanding social sycophancy in LLMs
di: Cheng, Myra, et al.
Pubblicazione: (2025)
di: Cheng, Myra, et al.
Pubblicazione: (2025)
From Flat to Structural: Enhancing Automated Short Answer Grading with GraphRAG
di: Chu, Yucheng, et al.
Pubblicazione: (2026)
di: Chu, Yucheng, et al.
Pubblicazione: (2026)
Exploring LGBTQ+ Bias in Generative AI Answers across Different Country and Religious Contexts
di: Vicsek, Lilla, et al.
Pubblicazione: (2024)
di: Vicsek, Lilla, et al.
Pubblicazione: (2024)
Classroom AI: Large Language Models as Grade-Specific Teachers
di: Oh, Jio, et al.
Pubblicazione: (2026)
di: Oh, Jio, et al.
Pubblicazione: (2026)
ASAG2024: A Combined Benchmark for Short Answer Grading
di: Meyer, Gérôme, et al.
Pubblicazione: (2024)
di: Meyer, Gérôme, et al.
Pubblicazione: (2024)
SteLLA: A Structured Grading System Using LLMs with RAG
di: Qiu, Hefei, et al.
Pubblicazione: (2025)
di: Qiu, Hefei, et al.
Pubblicazione: (2025)
"Sorry, I Didn't Catch That": How Speech Models Miss What Matters Most
di: Zhou, Kaitlyn, et al.
Pubblicazione: (2026)
di: Zhou, Kaitlyn, et al.
Pubblicazione: (2026)
Looks can be Deceptive: Distinguishing Repetition Disfluency from Reduplication
di: Ahmad, Arif, et al.
Pubblicazione: (2024)
di: Ahmad, Arif, et al.
Pubblicazione: (2024)
The Responsible Development of Automated Student Feedback with Generative AI
di: Lindsay, Euan D, et al.
Pubblicazione: (2023)
di: Lindsay, Euan D, et al.
Pubblicazione: (2023)
Using GPT-4 to Augment Unbalanced Data for Automatic Scoring
di: Fang, Luyang, et al.
Pubblicazione: (2023)
di: Fang, Luyang, et al.
Pubblicazione: (2023)
Artificial Intelligence Bias on English Language Learners in Automatic Scoring
di: Guo, Shuchen, et al.
Pubblicazione: (2025)
di: Guo, Shuchen, et al.
Pubblicazione: (2025)
Can LLMs Grade Short-Answer Reading Comprehension Questions : An Empirical Study with a Novel Dataset
di: Henkel, Owen, et al.
Pubblicazione: (2023)
di: Henkel, Owen, et al.
Pubblicazione: (2023)
REC-CBM: Rubric-Aware Error-Correction Concept Bottleneck Models for Trustworthy Open-Ended Grading
di: Zhao, Chengshuai, et al.
Pubblicazione: (2026)
di: Zhao, Chengshuai, et al.
Pubblicazione: (2026)
The Good, the Bad and the Constructive: Automatically Measuring Peer Review's Utility for Authors
di: Sadallah, Abdelrahman, et al.
Pubblicazione: (2025)
di: Sadallah, Abdelrahman, et al.
Pubblicazione: (2025)
Evaluating Reliability Asymmetries in Chinese Factual Search and AI Answers
di: Liu, Geng, et al.
Pubblicazione: (2025)
di: Liu, Geng, et al.
Pubblicazione: (2025)
"I Am the One and Only, Your Cyber BFF": Understanding the Impact of GenAI Requires Understanding the Impact of Anthropomorphic AI
di: Cheng, Myra, et al.
Pubblicazione: (2024)
di: Cheng, Myra, et al.
Pubblicazione: (2024)
Leveraging Lecture Content for Improved Feedback: Explorations with GPT-4 and Retrieval Augmented Generation
di: Jacobs, Sven, et al.
Pubblicazione: (2024)
di: Jacobs, Sven, et al.
Pubblicazione: (2024)
StereoDetect: Detecting Stereotypes and Anti-stereotypes the Correct Way Using Social Psychological Underpinnings
di: Shejole, Kaustubh Shivshankar, et al.
Pubblicazione: (2025)
di: Shejole, Kaustubh Shivshankar, et al.
Pubblicazione: (2025)
MedSimAI: Simulation and Formative Feedback Generation to Enhance Deliberate Practice in Medical Education
di: Hicke, Yann, et al.
Pubblicazione: (2025)
di: Hicke, Yann, et al.
Pubblicazione: (2025)
Open Source Language Models Can Provide Feedback: Evaluating LLMs' Ability to Help Students Using GPT-4-As-A-Judge
di: Koutcheme, Charles, et al.
Pubblicazione: (2024)
di: Koutcheme, Charles, et al.
Pubblicazione: (2024)
Intelligent Tutor: Leveraging ChatGPT and Microsoft Copilot Studio to Deliver a Generative AI Student Support and Feedback System within Teams
di: Chen, Wei-Yu
Pubblicazione: (2024)
di: Chen, Wei-Yu
Pubblicazione: (2024)
Unveiling the Invisible: Captioning Videos with Metaphors
di: Kalarani, Abisek Rajakumar, et al.
Pubblicazione: (2024)
di: Kalarani, Abisek Rajakumar, et al.
Pubblicazione: (2024)
Chain-of-Description: What I can understand, I can put into words
di: Guo, Jiaxin, et al.
Pubblicazione: (2025)
di: Guo, Jiaxin, et al.
Pubblicazione: (2025)
Statistical Comparative Analysis of Semantic Similarities and Model Transferability Across Datasets for Short Answer Grading
di: Bonthu, Sridevi, et al.
Pubblicazione: (2025)
di: Bonthu, Sridevi, et al.
Pubblicazione: (2025)
From tools to thieves: Measuring and understanding public perceptions of AI through crowdsourced metaphors
di: Cheng, Myra, et al.
Pubblicazione: (2025)
di: Cheng, Myra, et al.
Pubblicazione: (2025)
Emotion Entanglement and Bayesian Inference for Multi-Dimensional Emotion Understanding
di: Kotaprolu, Hemanth, et al.
Pubblicazione: (2026)
di: Kotaprolu, Hemanth, et al.
Pubblicazione: (2026)
What's color got to do with it? Face recognition in grayscale
di: Bhatta, Aman, et al.
Pubblicazione: (2023)
di: Bhatta, Aman, et al.
Pubblicazione: (2023)
Can Large Language Models Make the Grade? An Empirical Study Evaluating LLMs Ability to Mark Short Answer Questions in K-12 Education
di: Henkel, Owen, et al.
Pubblicazione: (2024)
di: Henkel, Owen, et al.
Pubblicazione: (2024)
Towards LLM-based Autograding for Short Textual Answers
di: Schneider, Johannes, et al.
Pubblicazione: (2023)
di: Schneider, Johannes, et al.
Pubblicazione: (2023)
ToxVidLM: A Multimodal Framework for Toxicity Detection in Code-Mixed Videos
di: Maity, Krishanu, et al.
Pubblicazione: (2024)
di: Maity, Krishanu, et al.
Pubblicazione: (2024)
GRAFT: A Graph-based Flow-aware Agentic Framework for Document-level Machine Translation
di: Dutta, Himanshu, et al.
Pubblicazione: (2025)
di: Dutta, Himanshu, et al.
Pubblicazione: (2025)
Understand the Implication: Learning to Think for Pragmatic Understanding
di: Sravanthi, Settaluri Lakshmi, et al.
Pubblicazione: (2025)
di: Sravanthi, Settaluri Lakshmi, et al.
Pubblicazione: (2025)
Documenti analoghi
-
How effective are VLMs in assisting humans in inferring the quality of mental models from Multimodal short answers?
di: Sil, Pritam, et al.
Pubblicazione: (2026) -
Can AI Assistance Aid in the Grading of Handwritten Answer Sheets?
di: Sil, Pritam, et al.
Pubblicazione: (2024) -
Recon, Answer, Verify: Agents in Search of Truth
di: Shukla, Satyam, et al.
Pubblicazione: (2025) -
I've got the "Answer"! Interpretation of LLMs Hidden States in Question Answering
di: Goloviznina, Valeriya, et al.
Pubblicazione: (2024) -
Can MLLMs generate human-like feedback in grading multimodal short answers?
di: Sil, Pritam, et al.
Pubblicazione: (2024)