Automated Long Answer Grading with RiceChem Dataset
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sonkar, Shashank, Ni, Kangqi, Lu, Lesa Tran, Kincaid, Kristi, Hutchinson, John S., Baraniuk, Richard G. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Pedagogical Alignment of Large Language Models
von: Sonkar, Shashank, et al.
Veröffentlicht: (2024)
von: Sonkar, Shashank, et al.
Veröffentlicht: (2024)
Many-Shot Regurgitation (MSR) Prompting
von: Sonkar, Shashank, et al.
Veröffentlicht: (2024)
von: Sonkar, Shashank, et al.
Veröffentlicht: (2024)
Marking: Visual Grading with Highlighting Errors and Annotating Missing Bits
von: Sonkar, Shashank, et al.
Veröffentlicht: (2024)
von: Sonkar, Shashank, et al.
Veröffentlicht: (2024)
CLEAR-3K: Assessing Causal Explanatory Capabilities in Language Models
von: Liu, Naiming, et al.
Veröffentlicht: (2025)
von: Liu, Naiming, et al.
Veröffentlicht: (2025)
Student Data Paradox and Curious Case of Single Student-Tutor Model: Regressive Side Effects of Training LLMs for Personalized Learning
von: Sonkar, Shashank, et al.
Veröffentlicht: (2024)
von: Sonkar, Shashank, et al.
Veröffentlicht: (2024)
Do LLMs Make Mistakes Like Students? Exploring Natural Alignment between Language Models and Human Error Patterns
von: Liu, Naiming, et al.
Veröffentlicht: (2025)
von: Liu, Naiming, et al.
Veröffentlicht: (2025)
MalAlgoQA: Pedagogical Evaluation of Counterfactual Reasoning in Large Language Models and Implications for AI in Education
von: Liu, Naiming, et al.
Veröffentlicht: (2024)
von: Liu, Naiming, et al.
Veröffentlicht: (2024)
LLM-based Cognitive Models of Students with Misconceptions
von: Sonkar, Shashank, et al.
Veröffentlicht: (2024)
von: Sonkar, Shashank, et al.
Veröffentlicht: (2024)
MetaCLASS: Metacognitive Coaching for Learning with Adaptive Self-regulation Support
von: Liu, Naiming, et al.
Veröffentlicht: (2026)
von: Liu, Naiming, et al.
Veröffentlicht: (2026)
When Can We Trust LLM Graders? Calibrating Confidence for Automated Assessment
von: Ferrer, Robinson, et al.
Veröffentlicht: (2026)
von: Ferrer, Robinson, et al.
Veröffentlicht: (2026)
FoundationalASSIST: An Educational Dataset for Foundational Knowledge Tracing and Pedagogical Grounding of LLMs
von: Worden, Eamon, et al.
Veröffentlicht: (2026)
von: Worden, Eamon, et al.
Veröffentlicht: (2026)
The Imitation Game for Educational AI
von: Sonkar, Shashank, et al.
Veröffentlicht: (2025)
von: Sonkar, Shashank, et al.
Veröffentlicht: (2025)
MalruleLib: Large-Scale Executable Misconception Reasoning with Step Traces for Modeling Student Thinking in Mathematics
von: Chen, Xinghe, et al.
Veröffentlicht: (2026)
von: Chen, Xinghe, et al.
Veröffentlicht: (2026)
Circuit Complexity of Hierarchical Knowledge Tracing and Implications for Log-Precision Transformers
von: Liu, Naiming, et al.
Veröffentlicht: (2026)
von: Liu, Naiming, et al.
Veröffentlicht: (2026)
Simulating Students or Sycophantic Problem Solving? On Misconception Faithfulness of LLM Simulators
von: Do, Heejin, et al.
Veröffentlicht: (2026)
von: Do, Heejin, et al.
Veröffentlicht: (2026)
Misconception Acquisition Dynamics in Large Language Models
von: Liu, Naiming, et al.
Veröffentlicht: (2026)
von: Liu, Naiming, et al.
Veröffentlicht: (2026)
Grade Guard: A Smart System for Short Answer Automated Grading
von: Dadu, Niharika, et al.
Veröffentlicht: (2025)
von: Dadu, Niharika, et al.
Veröffentlicht: (2025)
Focusing on Students, not Machines: Grounded Question Generation and Automated Answer Grading
von: Meyer, Gérôme, et al.
Veröffentlicht: (2025)
von: Meyer, Gérôme, et al.
Veröffentlicht: (2025)
Atomic Learning Objectives Labeling: A High-Resolution Approach for Physics Education
von: Liu, Naiming, et al.
Veröffentlicht: (2024)
von: Liu, Naiming, et al.
Veröffentlicht: (2024)
Enhancing Security and Strengthening Defenses in Automated Short-Answer Grading Systems
von: Yarmohammadtoosky, Sahar, et al.
Veröffentlicht: (2025)
von: Yarmohammadtoosky, Sahar, et al.
Veröffentlicht: (2025)
From Flat to Structural: Enhancing Automated Short Answer Grading with GraphRAG
von: Chu, Yucheng, et al.
Veröffentlicht: (2026)
von: Chu, Yucheng, et al.
Veröffentlicht: (2026)
How Long Reasoning Chains Influence LLMs' Judgment of Answer Factuality
von: Tu, Minzhu, et al.
Veröffentlicht: (2026)
von: Tu, Minzhu, et al.
Veröffentlicht: (2026)
Confidence Estimation in Automatic Short Answer Grading with LLMs
von: Cong, Longwei, et al.
Veröffentlicht: (2026)
von: Cong, Longwei, et al.
Veröffentlicht: (2026)
Scalable Generation and Validation of Isomorphic Physics Problems with GenAI
von: Liu, Naiming, et al.
Veröffentlicht: (2026)
von: Liu, Naiming, et al.
Veröffentlicht: (2026)
Synthetic Context Generation for Question Generation
von: Liu, Naiming, et al.
Veröffentlicht: (2024)
von: Liu, Naiming, et al.
Veröffentlicht: (2024)
Breaking the Mold: Nonlinear Ranking Function Synthesis Without Templates
von: Zhu, Shaowei, et al.
Veröffentlicht: (2024)
von: Zhu, Shaowei, et al.
Veröffentlicht: (2024)
Can LLMs Grade Short-Answer Reading Comprehension Questions : An Empirical Study with a Novel Dataset
von: Henkel, Owen, et al.
Veröffentlicht: (2023)
von: Henkel, Owen, et al.
Veröffentlicht: (2023)
Statistical Comparative Analysis of Semantic Similarities and Model Transferability Across Datasets for Short Answer Grading
von: Bonthu, Sridevi, et al.
Veröffentlicht: (2025)
von: Bonthu, Sridevi, et al.
Veröffentlicht: (2025)
Datasets for Multilingual Answer Sentence Selection
von: Gabburo, Matteo, et al.
Veröffentlicht: (2024)
von: Gabburo, Matteo, et al.
Veröffentlicht: (2024)
Enhancing LLM-Based Short Answer Grading with Retrieval-Augmented Generation
von: Chu, Yucheng, et al.
Veröffentlicht: (2025)
von: Chu, Yucheng, et al.
Veröffentlicht: (2025)
Estimating LLM Grading Ability and Response Difficulty in Automatic Short Answer Grading via Item Response Theory
von: Cong, Longwei, et al.
Veröffentlicht: (2026)
von: Cong, Longwei, et al.
Veröffentlicht: (2026)
Can LLMs Model Incorrect Student Reasoning? A Case Study on Distractor Generation
von: Zengaffinen, Yanick, et al.
Veröffentlicht: (2026)
von: Zengaffinen, Yanick, et al.
Veröffentlicht: (2026)
EduAdapt: A Question Answer Benchmark Dataset for Evaluating Grade-Level Adaptability in LLMs
von: Naeem, Numaan, et al.
Veröffentlicht: (2025)
von: Naeem, Numaan, et al.
Veröffentlicht: (2025)
CHiL(L)Grader: Calibrated Human-in-the-Loop Short-Answer Grading
von: Raikote, Pranav, et al.
Veröffentlicht: (2026)
von: Raikote, Pranav, et al.
Veröffentlicht: (2026)
Can GRPO Help LLMs Transcend Their Pretraining Origin?
von: Ni, Kangqi, et al.
Veröffentlicht: (2025)
von: Ni, Kangqi, et al.
Veröffentlicht: (2025)
Not All Synthetic Data Is Yours to Learn From
von: Alemohammad, Sina, et al.
Veröffentlicht: (2026)
von: Alemohammad, Sina, et al.
Veröffentlicht: (2026)
Training LLM-based Tutors to Improve Student Learning Outcomes in Dialogues
von: Scarlatos, Alexander, et al.
Veröffentlicht: (2025)
von: Scarlatos, Alexander, et al.
Veröffentlicht: (2025)
Modeling the Sacred: Considerations when Using Religious Texts in Natural Language Processing
von: Hutchinson, Ben
Veröffentlicht: (2024)
von: Hutchinson, Ben
Veröffentlicht: (2024)
BioACE: An Automated Framework for Biomedical Answer and Citation Evaluations
von: Gupta, Deepak, et al.
Veröffentlicht: (2026)
von: Gupta, Deepak, et al.
Veröffentlicht: (2026)
Enhancing Multi-Domain Automatic Short Answer Grading through an Explainable Neuro-Symbolic Pipeline
von: Künnecke, Felix, et al.
Veröffentlicht: (2024)
von: Künnecke, Felix, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Pedagogical Alignment of Large Language Models
von: Sonkar, Shashank, et al.
Veröffentlicht: (2024) -
Many-Shot Regurgitation (MSR) Prompting
von: Sonkar, Shashank, et al.
Veröffentlicht: (2024) -
Marking: Visual Grading with Highlighting Errors and Annotating Missing Bits
von: Sonkar, Shashank, et al.
Veröffentlicht: (2024) -
CLEAR-3K: Assessing Causal Explanatory Capabilities in Language Models
von: Liu, Naiming, et al.
Veröffentlicht: (2025) -
Student Data Paradox and Curious Case of Single Student-Tutor Model: Regressive Side Effects of Training LLMs for Personalized Learning
von: Sonkar, Shashank, et al.
Veröffentlicht: (2024)