Using Large Language Models to Assess Tutors' Performance in Reacting to Students Making Math Errors
Fuente:
arXiv
Saved in:
| Main Authors: | Kakarla, Sanjit, Thomas, Danielle, Lin, Jionghao, Gupta, Shivang, Koedinger, Kenneth R. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GPTutor: Great Personalized Tutor with Large Language Models for Personalized Learning Content Generation
by: Chen, Eason, et al.
Published: (2024)
by: Chen, Eason, et al.
Published: (2024)
Do Tutors Learn from Equity Training and Can Generative AI Assess It?
by: Thomas, Danielle R., et al.
Published: (2024)
by: Thomas, Danielle R., et al.
Published: (2024)
Comparing Few-Shot Prompting of GPT-4 LLMs with BERT Classifiers for Open-Response Assessment in Tutor Equity Training
by: Kakarla, Sanjit, et al.
Published: (2025)
by: Kakarla, Sanjit, et al.
Published: (2025)
Improving Assessment of Tutoring Practices using Retrieval-Augmented Generation
by: Han, Zifei FeiFei, et al.
Published: (2024)
by: Han, Zifei FeiFei, et al.
Published: (2024)
Leveraging LLMs to Assess Tutor Moves in Real-Life Dialogues: A Feasibility Study
by: Thomas, Danielle R., et al.
Published: (2025)
by: Thomas, Danielle R., et al.
Published: (2025)
VTutor for High-Impact Tutoring at Scale: Managing Engagement and Real-Time Multi-Screen Monitoring with P2P Connections
by: Chen, Eason, et al.
Published: (2025)
by: Chen, Eason, et al.
Published: (2025)
AI Knows Best? The Paradox of Expertise, AI-Reliance, and Performance in Educational Tutoring Decision-Making Tasks
by: Chen, Eason, et al.
Published: (2025)
by: Chen, Eason, et al.
Published: (2025)
Leveraging Large Language Models for Identifying Knowledge Components
by: Wang, Canwen, et al.
Published: (2025)
by: Wang, Canwen, et al.
Published: (2025)
How Can I Get It Right? Using GPT to Rephrase Incorrect Trainee Responses
by: Lin, Jionghao, et al.
Published: (2024)
by: Lin, Jionghao, et al.
Published: (2024)
Does Multiple Choice Have a Future in the Age of Generative AI? A Posttest-only RCT
by: Thomas, Danielle R., et al.
Published: (2024)
by: Thomas, Danielle R., et al.
Published: (2024)
Improving Automated Feedback Systems for Tutor Training in Low-Resource Scenarios through Data Augmentation
by: Xu, Chentianye, et al.
Published: (2025)
by: Xu, Chentianye, et al.
Published: (2025)
Toward Automated Qualitative Analysis: Leveraging Large Language Models for Tutoring Dialogue Evaluation
by: Gu, Megan, et al.
Published: (2025)
by: Gu, Megan, et al.
Published: (2025)
Combining Large Language Models with Tutoring System Intelligence: A Case Study in Caregiver Homework Support
by: Venugopalan, Devika, et al.
Published: (2024)
by: Venugopalan, Devika, et al.
Published: (2024)
Automatic Large Language Models Creation of Interactive Learning Lessons
by: Lin, Jionghao, et al.
Published: (2025)
by: Lin, Jionghao, et al.
Published: (2025)
Detecting LLM-Generated Short Answers and Effects on Learner Performance
by: Bhushan, Shambhavi, et al.
Published: (2025)
by: Bhushan, Shambhavi, et al.
Published: (2025)
Assessing the Impact and Underlying Pathways of Sequenced AI feedback on Student Learning
by: Cao, Jie, et al.
Published: (2026)
by: Cao, Jie, et al.
Published: (2026)
An Integrated Platform for Studying Learning with Intelligent Tutoring Systems: CTAT+TutorShop
by: Aleven, Vincent, et al.
Published: (2025)
by: Aleven, Vincent, et al.
Published: (2025)
LLM-based Multimodal Feedback Produces Equivalent Learning and Better Student Perceptions than Educator Feedback
by: Zhao, Chloe Qianhui, et al.
Published: (2026)
by: Zhao, Chloe Qianhui, et al.
Published: (2026)
Simulating Novice Students Using Machine Unlearning and Relearning in Large Language Models
by: Song, Jiajia, et al.
Published: (2026)
by: Song, Jiajia, et al.
Published: (2026)
SlideItRight: Using AI to Find Relevant Slides and Provide Feedback for Open-Ended Questions
by: Zhao, Chloe Qianhui, et al.
Published: (2025)
by: Zhao, Chloe Qianhui, et al.
Published: (2025)
Comparing RAG and GraphRAG for Page-Level Retrieval Question Answering on Math Textbook
by: Chen, Eason, et al.
Published: (2025)
by: Chen, Eason, et al.
Published: (2025)
Sakshm AI: Advancing AI-Assisted Coding Education for Engineering Students in India Through Socratic Tutoring and Comprehensive Feedback
by: Gupta, Raj, et al.
Published: (2025)
by: Gupta, Raj, et al.
Published: (2025)
Student Engagement with GenAI's Tutoring Feedback: A Mixed Methods Study
by: Jacobs, Sven, et al.
Published: (2025)
by: Jacobs, Sven, et al.
Published: (2025)
Aligning Tutor Discourse Supporting Rigorous Thinking with Tutee Content Mastery for Predicting Math Achievement
by: Abdelshiheed, Mark, et al.
Published: (2024)
by: Abdelshiheed, Mark, et al.
Published: (2024)
Beyond Final Answers: Evaluating Large Language Models for Math Tutoring
by: Gupta, Adit, et al.
Published: (2025)
by: Gupta, Adit, et al.
Published: (2025)
VTutor: An Animated Pedagogical Agent SDK that Provide Real Time Multi-Model Feedback
by: Chen, Eason, et al.
Published: (2025)
by: Chen, Eason, et al.
Published: (2025)
How Can I Improve? Using GPT to Highlight the Desired and Undesired Parts of Open-ended Responses
by: Lin, Jionghao, et al.
Published: (2024)
by: Lin, Jionghao, et al.
Published: (2024)
Deceptive, Disruptive, No Big Deal: Japanese People React to Simulated Dark Commercial Patterns
by: Seaborn, Katie, et al.
Published: (2024)
by: Seaborn, Katie, et al.
Published: (2024)
Does the TalkMoves Codebook Generalize to One-on-One Tutoring and Multimodal Interaction?
by: Focsan, Corina Luca, et al.
Published: (2026)
by: Focsan, Corina Luca, et al.
Published: (2026)
WIP: Large Language Model-Enhanced Smart Tutor for Undergraduate Circuit Analysis
by: Chen, Liangliang, et al.
Published: (2025)
by: Chen, Liangliang, et al.
Published: (2025)
The Future of Learning: Large Language Models through the Lens of Students
by: Zhang, He, et al.
Published: (2024)
by: Zhang, He, et al.
Published: (2024)
Intelligent Tutors for Adult Learners: An Analysis of Needs and Challenges
by: Gupta, Adit, et al.
Published: (2024)
by: Gupta, Adit, et al.
Published: (2024)
Beyond the AI Tutor: Social Learning with LLM Agents
by: Kumar, Harsh, et al.
Published: (2026)
by: Kumar, Harsh, et al.
Published: (2026)
From First Draft to Final Insight: A Multi-Agent Approach for Feedback Generation
by: Cao, Jie, et al.
Published: (2025)
by: Cao, Jie, et al.
Published: (2025)
ClickTree: A Tree-based Method for Predicting Math Students' Performance Based on Clickstream Data
by: Rohani, Narjes, et al.
Published: (2024)
by: Rohani, Narjes, et al.
Published: (2024)
LLMs to Support K-12 Teachers in Culturally Relevant Pedagogy: An AI Literacy Example
by: Wang, Jiayi, et al.
Published: (2025)
by: Wang, Jiayi, et al.
Published: (2025)
Enhancing LLM-Based Feedback: Insights from Intelligent Tutoring Systems and the Learning Sciences
by: Stamper, John, et al.
Published: (2024)
by: Stamper, John, et al.
Published: (2024)
The Missing Evaluation Axis: What 10,000 Student Submissions Reveal About AI Tutor Effectiveness
by: Niousha, Rose, et al.
Published: (2026)
by: Niousha, Rose, et al.
Published: (2026)
Do They Understand What They Are Using? -- Assessing Perception and Usage of Biometrics
by: Mecke, Lukas, et al.
Published: (2024)
by: Mecke, Lukas, et al.
Published: (2024)
Improving Hybrid Human-AI Tutoring by Differentiating Human Tutor Roles Based on Student Needs
by: Gurung, Ashish, et al.
Published: (2026)
by: Gurung, Ashish, et al.
Published: (2026)
Similar Items
-
GPTutor: Great Personalized Tutor with Large Language Models for Personalized Learning Content Generation
by: Chen, Eason, et al.
Published: (2024) -
Do Tutors Learn from Equity Training and Can Generative AI Assess It?
by: Thomas, Danielle R., et al.
Published: (2024) -
Comparing Few-Shot Prompting of GPT-4 LLMs with BERT Classifiers for Open-Response Assessment in Tutor Equity Training
by: Kakarla, Sanjit, et al.
Published: (2025) -
Improving Assessment of Tutoring Practices using Retrieval-Augmented Generation
by: Han, Zifei FeiFei, et al.
Published: (2024) -
Leveraging LLMs to Assess Tutor Moves in Real-Life Dialogues: A Feasibility Study
by: Thomas, Danielle R., et al.
Published: (2025)