Automated Feedback in Math Education: A Comparative Analysis of LLMs for Open-Ended Responses
Fuente:
arXiv
Salvato in:
| Autori principali: | Baral, Sami, Worden, Eamon, Lim, Wen-Chiang, Luo, Zhuang, Santorelli, Christopher, Gurung, Ashish, Heffernan, Neil |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
FoundationalASSIST: An Educational Dataset for Foundational Knowledge Tracing and Pedagogical Grounding of LLMs
di: Worden, Eamon, et al.
Pubblicazione: (2026)
di: Worden, Eamon, et al.
Pubblicazione: (2026)
DrawEduMath: Evaluating Vision Language Models with Expert-Annotated Students' Hand-Drawn Math Images
di: Baral, Sami, et al.
Pubblicazione: (2025)
di: Baral, Sami, et al.
Pubblicazione: (2025)
The Karp Dataset
di: DiCicco, Mason, et al.
Pubblicazione: (2025)
di: DiCicco, Mason, et al.
Pubblicazione: (2025)
Investigating the Robustness of Knowledge Tracing Models in the Presence of Student Concept Drift
di: Lee, Morgan, et al.
Pubblicazione: (2025)
di: Lee, Morgan, et al.
Pubblicazione: (2025)
Can Large Language Models Replicate ITS Feedback on Open-Ended Math Questions?
di: McNichols, Hunter, et al.
Pubblicazione: (2024)
di: McNichols, Hunter, et al.
Pubblicazione: (2024)
Pessimistic Verification for Open Ended Math Questions
di: Huang, Yanxing, et al.
Pubblicazione: (2025)
di: Huang, Yanxing, et al.
Pubblicazione: (2025)
Generating Planning Feedback for Open-Ended Programming Exercises with LLMs
di: Demirtaş, Mehmet Arif, et al.
Pubblicazione: (2025)
di: Demirtaş, Mehmet Arif, et al.
Pubblicazione: (2025)
The ASSISTments Ecosystem: Building a Platform That Brings Scientists and Teachers Together for Minimally Invasive Research on Human Learning and Teaching
di: Heffernan, Neil T., et al.
Pubblicazione: (2014)
di: Heffernan, Neil T., et al.
Pubblicazione: (2014)
Transparent Reference-free Automated Evaluation of Open-Ended User Survey Responses
di: An, Subin, et al.
Pubblicazione: (2025)
di: An, Subin, et al.
Pubblicazione: (2025)
Automate Knowledge Concept Tagging on Math Questions with LLMs
di: Li, Hang, et al.
Pubblicazione: (2024)
di: Li, Hang, et al.
Pubblicazione: (2024)
Automated Consistency Analysis of LLMs
di: Patwardhan, Aditya, et al.
Pubblicazione: (2025)
di: Patwardhan, Aditya, et al.
Pubblicazione: (2025)
Can LLMs $\textit{understand}$ Math? -- Exploring the Pitfalls in Mathematical Reasoning
di: Roy, Tiasa Singha, et al.
Pubblicazione: (2025)
di: Roy, Tiasa Singha, et al.
Pubblicazione: (2025)
Do LLMs Exhibit Human-Like Reasoning? Evaluating Theory of Mind in LLMs for Open-Ended Responses
di: Amirizaniani, Maryam, et al.
Pubblicazione: (2024)
di: Amirizaniani, Maryam, et al.
Pubblicazione: (2024)
Aspect-Based Sentiment Analysis for Open-Ended HR Survey Responses
di: Rink, Lois, et al.
Pubblicazione: (2024)
di: Rink, Lois, et al.
Pubblicazione: (2024)
ActivityNarrated: An Open-Ended Narrative Paradigm for Wearable Human Activity Understanding
di: Ray, Lala Shakti Swarup, et al.
Pubblicazione: (2026)
di: Ray, Lala Shakti Swarup, et al.
Pubblicazione: (2026)
Hard2Verify: A Step-Level Verification Benchmark for Open-Ended Frontier Math
di: Pandit, Shrey, et al.
Pubblicazione: (2025)
di: Pandit, Shrey, et al.
Pubblicazione: (2025)
MathHay: An Automated Benchmark for Long-Context Mathematical Reasoning in LLMs
di: Wang, Lei, et al.
Pubblicazione: (2024)
di: Wang, Lei, et al.
Pubblicazione: (2024)
A Comparative Evaluation of Structural Topic Models and BERTopic for Short, Open-Ended Survey Responses
di: Jiang, Yan, et al.
Pubblicazione: (2026)
di: Jiang, Yan, et al.
Pubblicazione: (2026)
Grading Open‐Ended Questions Using LLMs and RAG
di: Jacobo Farray Rodríguez, et al.
Pubblicazione: (2025)
di: Jacobo Farray Rodríguez, et al.
Pubblicazione: (2025)
Feedback Descent: Open-Ended Text Optimization via Pairwise Comparison
di: Lee, Yoonho, et al.
Pubblicazione: (2025)
di: Lee, Yoonho, et al.
Pubblicazione: (2025)
Concept Drift Detection for Knowledge Tracing
di: Morgan Lee, et al.
Pubblicazione: (2025)
di: Morgan Lee, et al.
Pubblicazione: (2025)
Structuring Open-Ended NAS: Semi-Automated Design Knowledge Structuring with LLMs for Efficient Neural Architecture Search
di: Sakuma, Yuiko, et al.
Pubblicazione: (2026)
di: Sakuma, Yuiko, et al.
Pubblicazione: (2026)
AutoLibra: Agent Metric Induction from Open-Ended Human Feedback
di: Zhu, Hao, et al.
Pubblicazione: (2025)
di: Zhu, Hao, et al.
Pubblicazione: (2025)
The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery
di: Lu, Chris, et al.
Pubblicazione: (2024)
di: Lu, Chris, et al.
Pubblicazione: (2024)
CHIRP: A Fine-Grained Benchmark for Open-Ended Response Evaluation in Vision-Language Models
di: Roger, Alexis, et al.
Pubblicazione: (2025)
di: Roger, Alexis, et al.
Pubblicazione: (2025)
A Multi-Agent Approach to Validate and Refine LLM-Generated Personalized Math Problems
di: Ikram, Fareya, et al.
Pubblicazione: (2026)
di: Ikram, Fareya, et al.
Pubblicazione: (2026)
On Creativity and Open-Endedness
di: Soros, L. B., et al.
Pubblicazione: (2024)
di: Soros, L. B., et al.
Pubblicazione: (2024)
The Hyperfitting Phenomenon: Sharpening and Stabilizing LLMs for Open-Ended Text Generation
di: Carlsson, Fredrik, et al.
Pubblicazione: (2024)
di: Carlsson, Fredrik, et al.
Pubblicazione: (2024)
Quality Diversity through Human Feedback: Towards Open-Ended Diversity-Driven Optimization
di: Ding, Li, et al.
Pubblicazione: (2023)
di: Ding, Li, et al.
Pubblicazione: (2023)
MIRROR: A Novel Approach for the Automated Evaluation of Open-Ended Question Generation
di: Deroy, Aniket, et al.
Pubblicazione: (2024)
di: Deroy, Aniket, et al.
Pubblicazione: (2024)
Open-Ended Multi-Modal Relational Reasoning for Video Question Answering
di: Luo, Haozheng, et al.
Pubblicazione: (2020)
di: Luo, Haozheng, et al.
Pubblicazione: (2020)
“What Are Some of the Things You Are Worried About?”: An Analysis of Youth's Open‐Ended Responses of Current Worries
di: Taylor Heffer, et al.
Pubblicazione: (2025)
di: Taylor Heffer, et al.
Pubblicazione: (2025)
Computation, Analysis and Preparation of Coastwide Oyster Population Data - Survey of Oyster Population and Associated Organisms
di: Heffernan, T. L.
Pubblicazione: ()
di: Heffernan, T. L.
Pubblicazione: ()
StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs
di: Jeune, Pierre Le, et al.
Pubblicazione: (2026)
di: Jeune, Pierre Le, et al.
Pubblicazione: (2026)
How Do Open‐Ended Interview and Closed‐Ended Questionnaire Responses Compare? A Matched Mixed‐Methods Study Examining Treatment Attitudes of Patients With Meniscal Tear and Persistent Pain
di: Paul M. Oh, et al.
Pubblicazione: (2026)
di: Paul M. Oh, et al.
Pubblicazione: (2026)
SlideItRight: Using AI to Find Relevant Slides and Provide Feedback for Open-Ended Questions
di: Zhao, Chloe Qianhui, et al.
Pubblicazione: (2025)
di: Zhao, Chloe Qianhui, et al.
Pubblicazione: (2025)
Representation Learning to Study Temporal Dynamics in Tutorial Scaffolding
di: Borchers, Conrad, et al.
Pubblicazione: (2026)
di: Borchers, Conrad, et al.
Pubblicazione: (2026)
AHP-Powered LLM Reasoning for Multi-Criteria Evaluation of Open-Ended Responses
di: Lu, Xiaotian, et al.
Pubblicazione: (2024)
di: Lu, Xiaotian, et al.
Pubblicazione: (2024)
Automated Distractor and Feedback Generation for Math Multiple-choice Questions via In-context Learning
di: McNichols, Hunter, et al.
Pubblicazione: (2023)
di: McNichols, Hunter, et al.
Pubblicazione: (2023)
MathConstraint: Automated Generation of Verified Combinatorial Reasoning Instances for LLMs
di: Pati, Viresh, et al.
Pubblicazione: (2026)
di: Pati, Viresh, et al.
Pubblicazione: (2026)
Documenti analoghi
-
FoundationalASSIST: An Educational Dataset for Foundational Knowledge Tracing and Pedagogical Grounding of LLMs
di: Worden, Eamon, et al.
Pubblicazione: (2026) -
DrawEduMath: Evaluating Vision Language Models with Expert-Annotated Students' Hand-Drawn Math Images
di: Baral, Sami, et al.
Pubblicazione: (2025) -
The Karp Dataset
di: DiCicco, Mason, et al.
Pubblicazione: (2025) -
Investigating the Robustness of Knowledge Tracing Models in the Presence of Student Concept Drift
di: Lee, Morgan, et al.
Pubblicazione: (2025) -
Can Large Language Models Replicate ITS Feedback on Open-Ended Math Questions?
di: McNichols, Hunter, et al.
Pubblicazione: (2024)