Unifying AI Tutor Evaluation: An Evaluation Taxonomy for Pedagogical Ability Assessment of LLM-Powered AI Tutors
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Maurya, Kaushal Kumar, Srivatsa, KV Aditya, Petukhova, Kseniia, Kochmar, Ekaterina |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Findings of the BEA 2025 Shared Task on Pedagogical Ability Assessment of AI-powered Tutors
von: Kochmar, Ekaterina, et al.
Veröffentlicht: (2025)
von: Kochmar, Ekaterina, et al.
Veröffentlicht: (2025)
Intent Matters: Enhancing AI Tutoring with Fine-Grained Pedagogical Intent Annotation
von: Petukhova, Kseniia, et al.
Veröffentlicht: (2025)
von: Petukhova, Kseniia, et al.
Veröffentlicht: (2025)
AITutor-EvalKit: Exploring the Capabilities of AI Tutors
von: Naeem, Numaan, et al.
Veröffentlicht: (2025)
von: Naeem, Numaan, et al.
Veröffentlicht: (2025)
Towards Reward Modeling for AI Tutors in Math Mistake Remediation
von: Petukhova, Kseniia, et al.
Veröffentlicht: (2026)
von: Petukhova, Kseniia, et al.
Veröffentlicht: (2026)
Pedagogy-driven Evaluation of Generative AI-powered Intelligent Tutoring Systems
von: Maurya, Kaushal Kumar, et al.
Veröffentlicht: (2025)
von: Maurya, Kaushal Kumar, et al.
Veröffentlicht: (2025)
Simulating LLM-to-LLM Tutoring for Multilingual Math Feedback
von: Tonga, Junior Cedric, et al.
Veröffentlicht: (2025)
von: Tonga, Junior Cedric, et al.
Veröffentlicht: (2025)
Harnessing the Power of Multiple Minds: Lessons Learned from LLM Routing
von: Srivatsa, KV Aditya, et al.
Veröffentlicht: (2024)
von: Srivatsa, KV Aditya, et al.
Veröffentlicht: (2024)
Can LLMs Reliably Simulate Real Students' Abilities in Mathematics and Reading Comprehension?
von: Srivatsa, KV Aditya, et al.
Veröffentlicht: (2025)
von: Srivatsa, KV Aditya, et al.
Veröffentlicht: (2025)
SelectLLM: Query-Aware Efficient Selection Algorithm for Large Language Models
von: Maurya, Kaushal Kumar, et al.
Veröffentlicht: (2024)
von: Maurya, Kaushal Kumar, et al.
Veröffentlicht: (2024)
LLMs cannot spot math errors, even when allowed to peek into the solution
von: Srivatsa, KV Aditya, et al.
Veröffentlicht: (2025)
von: Srivatsa, KV Aditya, et al.
Veröffentlicht: (2025)
What Makes Math Word Problems Challenging for LLMs?
von: Srivatsa, KV Aditya, et al.
Veröffentlicht: (2024)
von: Srivatsa, KV Aditya, et al.
Veröffentlicht: (2024)
A Fully Automated Pipeline for Conversational Discourse Annotation: Tree Scheme Generation and Labeling with Large Language Models
von: Petukhova, Kseniia, et al.
Veröffentlicht: (2025)
von: Petukhova, Kseniia, et al.
Veröffentlicht: (2025)
PetKaz at SemEval-2024 Task 8: Can Linguistics Capture the Specifics of LLM-generated Text?
von: Petukhova, Kseniia, et al.
Veröffentlicht: (2024)
von: Petukhova, Kseniia, et al.
Veröffentlicht: (2024)
PetKaz at SemEval-2024 Task 3: Advancing Emotion Classification with an LLM for Emotion-Cause Pair Extraction in Conversations
von: Kazakov, Roman, et al.
Veröffentlicht: (2024)
von: Kazakov, Roman, et al.
Veröffentlicht: (2024)
SafeTutors: Benchmarking Pedagogical Safety in AI Tutoring Systems
von: Hazra, Rima, et al.
Veröffentlicht: (2026)
von: Hazra, Rima, et al.
Veröffentlicht: (2026)
LLMs in Education: Novel Perspectives, Challenges, and Opportunities
von: Alhafni, Bashar, et al.
Veröffentlicht: (2024)
von: Alhafni, Bashar, et al.
Veröffentlicht: (2024)
Opportunities and Challenges of LLMs in Education: An NLP Perspective
von: Vajjala, Sowmya, et al.
Veröffentlicht: (2025)
von: Vajjala, Sowmya, et al.
Veröffentlicht: (2025)
GRADE: Generalizable Reasoning-Aware Dialogue Evaluation for AI Tutors
von: Bhalerao, Parth, et al.
Veröffentlicht: (2026)
von: Bhalerao, Parth, et al.
Veröffentlicht: (2026)
From Solver to Tutor: Evaluating the Pedagogical Intelligence of LLMs with KMP-Bench
von: Shi, Weikang, et al.
Veröffentlicht: (2026)
von: Shi, Weikang, et al.
Veröffentlicht: (2026)
MathTutorBench: A Benchmark for Measuring Open-ended Pedagogical Capabilities of LLM Tutors
von: Macina, Jakub, et al.
Veröffentlicht: (2025)
von: Macina, Jakub, et al.
Veröffentlicht: (2025)
Tutor Move Taxonomy: A Theory-Aligned Framework for Analyzing Instructional Moves in Tutoring
von: Zhou, Zhuqian, et al.
Veröffentlicht: (2026)
von: Zhou, Zhuqian, et al.
Veröffentlicht: (2026)
Evaluating the Impact of Advanced LLM Techniques on AI-Lecture Tutors for a Robotics Course
von: Kahl, Sebastian, et al.
Veröffentlicht: (2024)
von: Kahl, Sebastian, et al.
Veröffentlicht: (2024)
LLMs Are Already Good Tutors: Training-Free Prompt Optimization for Pedagogical Math Tutoring
von: Lee, Unggi, et al.
Veröffentlicht: (2026)
von: Lee, Unggi, et al.
Veröffentlicht: (2026)
BIPED: Pedagogically Informed Tutoring System for ESL Education
von: Kwon, Soonwoo, et al.
Veröffentlicht: (2024)
von: Kwon, Soonwoo, et al.
Veröffentlicht: (2024)
BD at BEA 2025 Shared Task: MPNet Ensembles for Pedagogical Mistake Identification and Localization in AI Tutor Responses
von: Rohan, Shadman, et al.
Veröffentlicht: (2025)
von: Rohan, Shadman, et al.
Veröffentlicht: (2025)
Listen, Correct, and Feed Back: Spoken Pedagogical Feedback Generation
von: Liang, Junhong, et al.
Veröffentlicht: (2026)
von: Liang, Junhong, et al.
Veröffentlicht: (2026)
PEARL: Training Socratic Tutors with Pedagogically Aligned Reinforcement Learning
von: Chang, Qikai, et al.
Veröffentlicht: (2026)
von: Chang, Qikai, et al.
Veröffentlicht: (2026)
Scaffolding Language Learning via Multi-modal Tutoring Systems with Pedagogical Instructions
von: Liu, Zhengyuan, et al.
Veröffentlicht: (2024)
von: Liu, Zhengyuan, et al.
Veröffentlicht: (2024)
UNVEILING: What Makes Linguistics Olympiad Puzzles Tricky for LLMs?
von: Choudhary, Mukund, et al.
Veröffentlicht: (2025)
von: Choudhary, Mukund, et al.
Veröffentlicht: (2025)
MMTutorBench: The First Multimodal Benchmark for AI Math Tutoring
von: Yang, Tengchao, et al.
Veröffentlicht: (2025)
von: Yang, Tengchao, et al.
Veröffentlicht: (2025)
DeepTutor: Towards Agentic Personalized Tutoring
von: Zhao, Bingxi, et al.
Veröffentlicht: (2026)
von: Zhao, Bingxi, et al.
Veröffentlicht: (2026)
REFeREE: A REference-FREE Model-Based Metric for Text Simplification
von: Huang, Yichen, et al.
Veröffentlicht: (2024)
von: Huang, Yichen, et al.
Veröffentlicht: (2024)
Generative AI and Its Impact on Personalized Intelligent Tutoring Systems
von: Maity, Subhankar, et al.
Veröffentlicht: (2024)
von: Maity, Subhankar, et al.
Veröffentlicht: (2024)
"Would You Want an AI Tutor?" Understanding Stakeholder Perceptions of LLM-based Systems in the Classroom
von: Fuligni, Caterina, et al.
Veröffentlicht: (2025)
von: Fuligni, Caterina, et al.
Veröffentlicht: (2025)
Assessing the Pedagogical Readiness of Large Language Models as AI Tutors in Low-Resource Contexts: A Case Study of Nepal's K-10 Curriculum
von: Acharya, Pratyush, et al.
Veröffentlicht: (2026)
von: Acharya, Pratyush, et al.
Veröffentlicht: (2026)
Teaching Through Analogies: A Modular Pipeline for Educational Analogy Generation
von: Barakat, Mariam, et al.
Veröffentlicht: (2026)
von: Barakat, Mariam, et al.
Veröffentlicht: (2026)
Large Language Models Approach Expert Pedagogical Quality in Math Tutoring but Differ in Instructional and Linguistic Profiles
von: Abdulsalam, Ramatu Oiza, et al.
Veröffentlicht: (2025)
von: Abdulsalam, Ramatu Oiza, et al.
Veröffentlicht: (2025)
How Teachers Can Use Large Language Models and Bloom's Taxonomy to Create Educational Quizzes
von: Elkins, Sabina, et al.
Veröffentlicht: (2024)
von: Elkins, Sabina, et al.
Veröffentlicht: (2024)
Beyond the AI Tutor: Social Learning with LLM Agents
von: Kumar, Harsh, et al.
Veröffentlicht: (2026)
von: Kumar, Harsh, et al.
Veröffentlicht: (2026)
CourseAssist: Pedagogically Appropriate AI Tutor for Computer Science Education
von: Feng, Ty, et al.
Veröffentlicht: (2024)
von: Feng, Ty, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Findings of the BEA 2025 Shared Task on Pedagogical Ability Assessment of AI-powered Tutors
von: Kochmar, Ekaterina, et al.
Veröffentlicht: (2025) -
Intent Matters: Enhancing AI Tutoring with Fine-Grained Pedagogical Intent Annotation
von: Petukhova, Kseniia, et al.
Veröffentlicht: (2025) -
AITutor-EvalKit: Exploring the Capabilities of AI Tutors
von: Naeem, Numaan, et al.
Veröffentlicht: (2025) -
Towards Reward Modeling for AI Tutors in Math Mistake Remediation
von: Petukhova, Kseniia, et al.
Veröffentlicht: (2026) -
Pedagogy-driven Evaluation of Generative AI-powered Intelligent Tutoring Systems
von: Maurya, Kaushal Kumar, et al.
Veröffentlicht: (2025)