Comparing Few-Shot Prompting of GPT-4 LLMs with BERT Classifiers for Open-Response Assessment in Tutor Equity Training
Fuente:
arXiv
Saved in:
| Main Authors: | Kakarla, Sanjit, Borchers, Conrad, Thomas, Danielle, Bhushan, Shambhavi, Koedinger, Kenneth R. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Do Tutors Learn from Equity Training and Can Generative AI Assess It?
by: Thomas, Danielle R., et al.
Published: (2024)
by: Thomas, Danielle R., et al.
Published: (2024)
Does Multiple Choice Have a Future in the Age of Generative AI? A Posttest-only RCT
by: Thomas, Danielle R., et al.
Published: (2024)
by: Thomas, Danielle R., et al.
Published: (2024)
Using Large Language Models to Assess Tutors' Performance in Reacting to Students Making Math Errors
by: Kakarla, Sanjit, et al.
Published: (2024)
by: Kakarla, Sanjit, et al.
Published: (2024)
Detecting LLM-Generated Short Answers and Effects on Learner Performance
by: Bhushan, Shambhavi, et al.
Published: (2025)
by: Bhushan, Shambhavi, et al.
Published: (2025)
Evaluating a Data-Driven Redesign Process for Intelligent Tutoring Systems
by: Lyu, Qianru, et al.
Published: (2026)
by: Lyu, Qianru, et al.
Published: (2026)
VTutor for High-Impact Tutoring at Scale: Managing Engagement and Real-Time Multi-Screen Monitoring with P2P Connections
by: Chen, Eason, et al.
Published: (2025)
by: Chen, Eason, et al.
Published: (2025)
Improving Assessment of Tutoring Practices using Retrieval-Augmented Generation
by: Han, Zifei FeiFei, et al.
Published: (2024)
by: Han, Zifei FeiFei, et al.
Published: (2024)
An Integrated Platform for Studying Learning with Intelligent Tutoring Systems: CTAT+TutorShop
by: Aleven, Vincent, et al.
Published: (2025)
by: Aleven, Vincent, et al.
Published: (2025)
Leveraging LLMs to Assess Tutor Moves in Real-Life Dialogues: A Feasibility Study
by: Thomas, Danielle R., et al.
Published: (2025)
by: Thomas, Danielle R., et al.
Published: (2025)
Improving Automated Feedback Systems for Tutor Training in Low-Resource Scenarios through Data Augmentation
by: Xu, Chentianye, et al.
Published: (2025)
by: Xu, Chentianye, et al.
Published: (2025)
How Learner Control and Explainable Learning Analytics on Skill Mastery Shape Student Desires to Finish and Avoid Loss in Tutored Practice
by: Borchers, Conrad, et al.
Published: (2024)
by: Borchers, Conrad, et al.
Published: (2024)
Not Everyone Wins with LLMs: Behavioral Patterns and Pedagogical Implications for AI Literacy in Programmatic Data Science
by: Ma, Qianou, et al.
Published: (2025)
by: Ma, Qianou, et al.
Published: (2025)
GPTutor: Great Personalized Tutor with Large Language Models for Personalized Learning Content Generation
by: Chen, Eason, et al.
Published: (2024)
by: Chen, Eason, et al.
Published: (2024)
How Can I Improve? Using GPT to Highlight the Desired and Undesired Parts of Open-ended Responses
by: Lin, Jionghao, et al.
Published: (2024)
by: Lin, Jionghao, et al.
Published: (2024)
Does the TalkMoves Codebook Generalize to One-on-One Tutoring and Multimodal Interaction?
by: Focsan, Corina Luca, et al.
Published: (2026)
by: Focsan, Corina Luca, et al.
Published: (2026)
AI Knows Best? The Paradox of Expertise, AI-Reliance, and Performance in Educational Tutoring Decision-Making Tasks
by: Chen, Eason, et al.
Published: (2025)
by: Chen, Eason, et al.
Published: (2025)
How Can I Get It Right? Using GPT to Rephrase Incorrect Trainee Responses
by: Lin, Jionghao, et al.
Published: (2024)
by: Lin, Jionghao, et al.
Published: (2024)
What Should We Engineer in Prompts? Training Humans in Requirement-Driven LLM Use
by: Ma, Qianou, et al.
Published: (2024)
by: Ma, Qianou, et al.
Published: (2024)
Combining Large Language Models with Tutoring System Intelligence: A Case Study in Caregiver Homework Support
by: Venugopalan, Devika, et al.
Published: (2024)
by: Venugopalan, Devika, et al.
Published: (2024)
How to Teach Programming in the AI Era? Using LLMs as a Teachable Agent for Debugging
by: Ma, Qianou, et al.
Published: (2023)
by: Ma, Qianou, et al.
Published: (2023)
Toward Automated Qualitative Analysis: Leveraging Large Language Models for Tutoring Dialogue Evaluation
by: Gu, Megan, et al.
Published: (2025)
by: Gu, Megan, et al.
Published: (2025)
SlideItRight: Using AI to Find Relevant Slides and Provide Feedback for Open-Ended Questions
by: Zhao, Chloe Qianhui, et al.
Published: (2025)
by: Zhao, Chloe Qianhui, et al.
Published: (2025)
AI2T: Building Trustable AI Tutors by Interactively Teaching a Self-Aware Learning Agent
by: Weitekamp, Daniel, et al.
Published: (2024)
by: Weitekamp, Daniel, et al.
Published: (2024)
Combining Log Data and Collaborative Dialogue Features to Predict Project Quality in Middle School AI Education
by: Borchers, Conrad, et al.
Published: (2025)
by: Borchers, Conrad, et al.
Published: (2025)
Understanding Student Effort Using Response-Time Propensities During Problem Solving
by: Borchers, Conrad, et al.
Published: (2026)
by: Borchers, Conrad, et al.
Published: (2026)
Toward Trait-Aware Learning Analytics
by: Borchers, Conrad, et al.
Published: (2026)
by: Borchers, Conrad, et al.
Published: (2026)
Simulating Learners' Task-Selection Strategies and System Constraints in Mastery Learning
by: Noh, Haley, et al.
Published: (2026)
by: Noh, Haley, et al.
Published: (2026)
Leveraging Large Language Models for Identifying Knowledge Components
by: Wang, Canwen, et al.
Published: (2025)
by: Wang, Canwen, et al.
Published: (2025)
Visualizing Intelligent Tutor Interactions for Responsive Pedagogy
by: Guo, Grace, et al.
Published: (2024)
by: Guo, Grace, et al.
Published: (2024)
TutorUp: What If Your Students Were Simulated? Training Tutors to Address Engagement Challenges in Online Learning
by: Pan, Sitong, et al.
Published: (2025)
by: Pan, Sitong, et al.
Published: (2025)
End User Authoring of Personalized Content Classifiers: Comparing Example Labeling, Rule Writing, and LLM Prompting
by: Wang, Leijie, et al.
Published: (2024)
by: Wang, Leijie, et al.
Published: (2024)
Cultural Prompting Improves the Empathy and Cultural Responsiveness of GPT-Generated Therapy Responses
by: Xie, Serena Jinchen, et al.
Published: (2025)
by: Xie, Serena Jinchen, et al.
Published: (2025)
Navigating Equity and Reflexive Practices in Gigwork Design: A Journey Mapping Experience
by: Boyd, Alicia E., et al.
Published: (2025)
by: Boyd, Alicia E., et al.
Published: (2025)
Facial-Expression-Aware Prompting for Empathetic LLM Tutoring
by: Feng, Shuangquan, et al.
Published: (2026)
by: Feng, Shuangquan, et al.
Published: (2026)
LLM-based Multimodal Feedback Produces Equivalent Learning and Better Student Perceptions than Educator Feedback
by: Zhao, Chloe Qianhui, et al.
Published: (2026)
by: Zhao, Chloe Qianhui, et al.
Published: (2026)
Comparing RAG and GraphRAG for Page-Level Retrieval Question Answering on Math Textbook
by: Chen, Eason, et al.
Published: (2025)
by: Chen, Eason, et al.
Published: (2025)
Learning to Use AI for Learning: Teaching Responsible Use of AI Chatbot to K-12 Students Through an AI Literacy Module
by: Xiao, Ruiwei, et al.
Published: (2025)
by: Xiao, Ruiwei, et al.
Published: (2025)
VTutor: An Animated Pedagogical Agent SDK that Provide Real Time Multi-Model Feedback
by: Chen, Eason, et al.
Published: (2025)
by: Chen, Eason, et al.
Published: (2025)
ActiveAI: Enabling K-12 AI Literacy Education & Analytics at Scale
by: Xiao, Ruiwei, et al.
Published: (2024)
by: Xiao, Ruiwei, et al.
Published: (2024)
Augmenting Human-Annotated Training Data with Large Language Model Generation and Distillation in Open-Response Assessment
by: Borchers, Conrad, et al.
Published: (2025)
by: Borchers, Conrad, et al.
Published: (2025)
Similar Items
-
Do Tutors Learn from Equity Training and Can Generative AI Assess It?
by: Thomas, Danielle R., et al.
Published: (2024) -
Does Multiple Choice Have a Future in the Age of Generative AI? A Posttest-only RCT
by: Thomas, Danielle R., et al.
Published: (2024) -
Using Large Language Models to Assess Tutors' Performance in Reacting to Students Making Math Errors
by: Kakarla, Sanjit, et al.
Published: (2024) -
Detecting LLM-Generated Short Answers and Effects on Learner Performance
by: Bhushan, Shambhavi, et al.
Published: (2025) -
Evaluating a Data-Driven Redesign Process for Intelligent Tutoring Systems
by: Lyu, Qianru, et al.
Published: (2026)