Utility-Preserving De-Identification for Math Tutoring: Investigating Numeric Ambiguity in the MathEd-PII Benchmark Dataset
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Zhuqian, Vanacore, Kirk, Ahtisham, Bakhtawar, Lee, Jinsook, Pietrzak, Doug, Hedley, Daryl, Dias, Jorge, Shaw, Chris, Schäfer, Ruth, Kizilcec, René F. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AI Annotation Orchestration: Evaluating LLM verifiers to Improve the Quality of LLM Annotations in Learning Analytics
by: Ahtisham, Bakhtawar, et al.
Published: (2025)
by: Ahtisham, Bakhtawar, et al.
Published: (2025)
Domain-Adapted Retrieval for In-Context Annotation of Pedagogical Dialogue Acts
by: Lee, Jinsook, et al.
Published: (2026)
by: Lee, Jinsook, et al.
Published: (2026)
LLM Reasoning Predicts When Models Are Right: Evidence from Coding Classroom Discourse
by: Ahtisham, Bakhtawar, et al.
Published: (2026)
by: Ahtisham, Bakhtawar, et al.
Published: (2026)
Codebook-Injected Dialogue Segmentation for Multi-Utterance Constructs Annotation: LLM-Assisted and Gold-Label-Free Evaluation
by: Lee, Jinsook, et al.
Published: (2026)
by: Lee, Jinsook, et al.
Published: (2026)
Million Tutoring Moves (MTM): An Open Multimodal Dataset for the Science of Tutoring
by: Kizilcec, René, et al.
Published: (2026)
by: Kizilcec, René, et al.
Published: (2026)
Optimizing LLM Annotation of Classroom Discourse through Multi-Agent Orchestration
by: Ahtisham, Bakhtawar, et al.
Published: (2026)
by: Ahtisham, Bakhtawar, et al.
Published: (2026)
Sandpiper: Orchestrated AI-Annotation for Educational Discourse at Scale
by: Hedley, Daryl, et al.
Published: (2026)
by: Hedley, Daryl, et al.
Published: (2026)
Tutor Move Taxonomy: A Theory-Aligned Framework for Analyzing Instructional Moves in Tutoring
by: Zhou, Zhuqian, et al.
Published: (2026)
by: Zhou, Zhuqian, et al.
Published: (2026)
How well do Large Language Models Recognize Instructional Moves? Establishing Baselines for Foundation Models in Educational Discourse
by: Vanacore, Kirk, et al.
Published: (2025)
by: Vanacore, Kirk, et al.
Published: (2025)
MathBuddy: A Multimodal System for Affective Math Tutoring
by: Kar, Debanjana, et al.
Published: (2025)
by: Kar, Debanjana, et al.
Published: (2025)
Poor Alignment and Steerability of Large Language Models: Evidence from College Admission Essays
by: Lee, Jinsook, et al.
Published: (2025)
by: Lee, Jinsook, et al.
Published: (2025)
Effective and Scalable Math Support: Evidence on the Impact of an AI- Tutor on Math Achievement in Ghana
by: Henkel, Owen, et al.
Published: (2024)
by: Henkel, Owen, et al.
Published: (2024)
The Path to Conversational AI Tutors: Integrating Tutoring Best Practices and Targeted Technologies to Produce Scalable AI Agents
by: Vanacore, Kirk, et al.
Published: (2026)
by: Vanacore, Kirk, et al.
Published: (2026)
Does the TalkMoves Codebook Generalize to One-on-One Tutoring and Multimodal Interaction?
by: Focsan, Corina Luca, et al.
Published: (2026)
by: Focsan, Corina Luca, et al.
Published: (2026)
Author Intent: Eliminating Ambiguity in MathML
by: Carlisle, David, et al.
Published: (2024)
by: Carlisle, David, et al.
Published: (2024)
Simulating LLM-to-LLM Tutoring for Multilingual Math Feedback
by: Tonga, Junior Cedric, et al.
Published: (2025)
by: Tonga, Junior Cedric, et al.
Published: (2025)
Modernizing Ground Truth: Four Shifts Toward Improving Reliability and Validity in AI in Education
by: Thomas, Danielle R., et al.
Published: (2026)
by: Thomas, Danielle R., et al.
Published: (2026)
LLMs Are Already Good Tutors: Training-Free Prompt Optimization for Pedagogical Math Tutoring
by: Lee, Unggi, et al.
Published: (2026)
by: Lee, Unggi, et al.
Published: (2026)
The Digital Divide in Generative AI: Evidence from Large Language Model Use in College Admissions Essays
by: Lee, Jinsook, et al.
Published: (2026)
by: Lee, Jinsook, et al.
Published: (2026)
The Life Cycle of Large Language Models: A Review of Biases in Education
by: Lee, Jinsook, et al.
Published: (2024)
by: Lee, Jinsook, et al.
Published: (2024)
The life cycle of large language models in education: A framework for understanding sources of bias
by: Jinsook Lee, et al.
Published: (2024)
by: Jinsook Lee, et al.
Published: (2024)
MMTutorBench: The First Multimodal Benchmark for AI Math Tutoring
by: Yang, Tengchao, et al.
Published: (2025)
by: Yang, Tengchao, et al.
Published: (2025)
AI-Powered Math Tutoring: Platform for Personalized and Adaptive Education
by: Chudziak, Jarosław A., et al.
Published: (2025)
by: Chudziak, Jarosław A., et al.
Published: (2025)
Towards Reward Modeling for AI Tutors in Math Mistake Remediation
by: Petukhova, Kseniia, et al.
Published: (2026)
by: Petukhova, Kseniia, et al.
Published: (2026)
MathTutorBench: A Benchmark for Measuring Open-ended Pedagogical Capabilities of LLM Tutors
by: Macina, Jakub, et al.
Published: (2025)
by: Macina, Jakub, et al.
Published: (2025)
Shiksha Copilot: Teacher-AI Collaboration for Curating and Customizing Lesson Plans in Low-Resource Schools
by: Dennison, Deepak Varuvel, et al.
Published: (2025)
by: Dennison, Deepak Varuvel, et al.
Published: (2025)
Beyond Final Answers: Evaluating Large Language Models for Math Tutoring
by: Gupta, Adit, et al.
Published: (2025)
by: Gupta, Adit, et al.
Published: (2025)
Real and Complex Analysis: Solutions to Problems in Amer. Math. Monthly, Math. Magazine, College Math. J., Elemente der Math., Crux Math., EMS Newsletter, Math. Gazette
by: Mortini, Raymond
Published: (2025)
by: Mortini, Raymond
Published: (2025)
Evaluating the Efficacy of an Intelligent Tutoring System That Integrates Affective Supports Into Math Learning
by: Mingyu Feng, et al.
Published: (2025)
by: Mingyu Feng, et al.
Published: (2025)
Algorithms for College Admissions Decision Support: Impacts of Policy Change and Inherent Variability
by: Lee, Jinsook, et al.
Published: (2024)
by: Lee, Jinsook, et al.
Published: (2024)
MegaMath: Pushing the Limits of Open Math Corpora
by: Zhou, Fan, et al.
Published: (2025)
by: Zhou, Fan, et al.
Published: (2025)
MathArena: Evaluating LLMs on Uncontaminated Math Competitions
by: Balunović, Mislav, et al.
Published: (2025)
by: Balunović, Mislav, et al.
Published: (2025)
Math Anxiety, Math Avoidance, Participation in Math: A Summary of Research Curriculum, Recommendations for Career Education.
by: Junghans, Barbara J.
Published: (1980)
by: Junghans, Barbara J.
Published: (1980)
AppliedMath
Published: (2022)
Published: (2022)
Maths & Music
by: Winterson, Julia
Published: (2024)
by: Winterson, Julia
Published: (2024)
IndoMath
Published: (2020)
Published: (2020)
Tangible Math
by: Scarlatos, Lori L.
Published: (2006)
by: Scarlatos, Lori L.
Published: (2006)
Where's the Math?
Published: (2003)
Published: (2003)
Math in the Library?
by: Henry, Robin
Published: (2004)
by: Henry, Robin
Published: (2004)
Aligning Tutor Discourse Supporting Rigorous Thinking with Tutee Content Mastery for Predicting Math Achievement
by: Abdelshiheed, Mark, et al.
Published: (2024)
by: Abdelshiheed, Mark, et al.
Published: (2024)
Similar Items
-
AI Annotation Orchestration: Evaluating LLM verifiers to Improve the Quality of LLM Annotations in Learning Analytics
by: Ahtisham, Bakhtawar, et al.
Published: (2025) -
Domain-Adapted Retrieval for In-Context Annotation of Pedagogical Dialogue Acts
by: Lee, Jinsook, et al.
Published: (2026) -
LLM Reasoning Predicts When Models Are Right: Evidence from Coding Classroom Discourse
by: Ahtisham, Bakhtawar, et al.
Published: (2026) -
Codebook-Injected Dialogue Segmentation for Multi-Utterance Constructs Annotation: LLM-Assisted and Gold-Label-Free Evaluation
by: Lee, Jinsook, et al.
Published: (2026) -
Million Tutoring Moves (MTM): An Open Multimodal Dataset for the Science of Tutoring
by: Kizilcec, René, et al.
Published: (2026)