Assessing Student Errors in Experimentation Using Artificial Intelligence and Large Language Models: A Comparative Study with Human Raters
Fuente:
arXiv
Saved in:
| Main Authors: | Bewersdorff, Arne, Seßler, Kathrin, Baur, Armin, Kasneci, Enkelejda, Nerdel, Claudia |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Adaptive Feedback with AI: Comparing the Feedback Quality of LLMs and Teachers on Experimentation Protocols
by: Seßler, Kathrin, et al.
Published: (2025)
by: Seßler, Kathrin, et al.
Published: (2025)
Taking the Next Step with Generative Artificial Intelligence: The Transformative Role of Multimodal Large Language Models in Science Education
by: Bewersdorff, Arne, et al.
Published: (2024)
by: Bewersdorff, Arne, et al.
Published: (2024)
Benchmarking Large Language Models for Math Reasoning Tasks
by: Seßler, Kathrin, et al.
Published: (2024)
by: Seßler, Kathrin, et al.
Published: (2024)
Assessing Professional Development on Experimentation as a Method of Inquiry‐Based Science Teaching Framework and Principal Results
by: Markus Emden, et al.
Published: (2025)
by: Markus Emden, et al.
Published: (2025)
Can AI grade your essays? A comparative analysis of large language models and teacher ratings in multidimensional essay scoring
by: Seßler, Kathrin, et al.
Published: (2024)
by: Seßler, Kathrin, et al.
Published: (2024)
Enriching Tabular Data with Contextual LLM Embeddings: A Comprehensive Ablation Study for Ensemble Classifiers
by: Kasneci, Gjergji, et al.
Published: (2024)
by: Kasneci, Gjergji, et al.
Published: (2024)
Position: Uncertainty Quantification Needs Reassessment for Large-language Model Agents
by: Kirchhof, Michael, et al.
Published: (2025)
by: Kirchhof, Michael, et al.
Published: (2025)
Sycophancy is an Educational Safety Risk: Why LLM Tutors Need Sycophancy Benchmarks
by: Kasneci, Enkelejda, et al.
Published: (2026)
by: Kasneci, Enkelejda, et al.
Published: (2026)
Exploring User Acceptance and Concerns toward LLM-powered Conversational Agents in Immersive Extended Reality
by: Bozkir, Efe, et al.
Published: (2025)
by: Bozkir, Efe, et al.
Published: (2025)
Lazy or Efficient? Towards Accessible Eye-Tracking Event Detection Using LLMs
by: Guo, Dongyang, et al.
Published: (2026)
by: Guo, Dongyang, et al.
Published: (2026)
Evaluating Usability and Engagement of Large Language Models in Virtual Reality for Traditional Scottish Curling
by: Lau, Ka Hei Carrie, et al.
Published: (2024)
by: Lau, Ka Hei Carrie, et al.
Published: (2024)
DART: An Automated End-to-End Object Detection Pipeline with Data Diversification, Open-Vocabulary Bounding Box Annotation, Pseudo-Label Review, and Model Training
by: Xin, Chen, et al.
Published: (2024)
by: Xin, Chen, et al.
Published: (2024)
Stepwise Self-Consistent Mathematical Reasoning with Large Language Models
by: Zhao, Zilong, et al.
Published: (2024)
by: Zhao, Zilong, et al.
Published: (2024)
MemeScouts@LT-EDI 2026: Asking the Right Questions -- Prompted Weak Supervision for Meme Hate Speech Detection
by: Bueno, Ivo, et al.
Published: (2026)
by: Bueno, Ivo, et al.
Published: (2026)
Towards Human-centered Explainable AI: A Survey of User Studies for Model Explanations
by: Rong, Yao, et al.
Published: (2022)
by: Rong, Yao, et al.
Published: (2022)
User Intent Recognition and Satisfaction with Large Language Models: A User Study with ChatGPT
by: Bodonhelyi, Anna, et al.
Published: (2024)
by: Bodonhelyi, Anna, et al.
Published: (2024)
Simulating Validity: Modal Decoupling in MLLM Generated Feedback on Science Drawings
by: Bewersdorff, Arne, et al.
Published: (2026)
by: Bewersdorff, Arne, et al.
Published: (2026)
Untersuchung der Effektivität zweier Fortbildungsformate zum Experimentieren mit dem Fokus auf das Unterrichtshandeln
by: Bewersdorff, Arne
Published: (2025)
by: Bewersdorff, Arne
Published: (2025)
Faithful Attention Explainer: Verbalizing Decisions Based on Discriminative Features
by: Rong, Yao, et al.
Published: (2024)
by: Rong, Yao, et al.
Published: (2024)
Embedding Large Language Models into Extended Reality: Opportunities and Challenges for Inclusion, Engagement, and Privacy
by: Bozkir, Efe, et al.
Published: (2024)
by: Bozkir, Efe, et al.
Published: (2024)
Automated Visual Attention Detection using Mobile Eye Tracking in Behavioral Classroom Studies
by: Bozkir, Efe, et al.
Published: (2025)
by: Bozkir, Efe, et al.
Published: (2025)
I-CEE: Tailoring Explanations of Image Classification Models to User Expertise
by: Rong, Yao, et al.
Published: (2023)
by: Rong, Yao, et al.
Published: (2023)
Using Large Language Models to Assess Tutors' Performance in Reacting to Students Making Math Errors
by: Kakarla, Sanjit, et al.
Published: (2024)
by: Kakarla, Sanjit, et al.
Published: (2024)
The Use of Artificial Intelligence Tools in Assessing Content Validity: A Comparative Study with Human Experts
by: Gurdil, Hatice, et al.
Published: (2025)
by: Gurdil, Hatice, et al.
Published: (2025)
Zero-Shot Segmentation of Eye Features Using the Segment Anything Model (SAM)
by: Maquiling, Virmarie, et al.
Published: (2023)
by: Maquiling, Virmarie, et al.
Published: (2023)
Conversational AI as a Catalyst for Informal Learning: An Empirical Large-Scale Study on LLM Use in Everyday Learning
by: Terzimehić, Nađa, et al.
Published: (2025)
by: Terzimehić, Nađa, et al.
Published: (2025)
Multimodal Behavioral Patterns Analysis with Eye-Tracking and LLM-Based Reasoning
by: Guo, Dongyang, et al.
Published: (2025)
by: Guo, Dongyang, et al.
Published: (2025)
Exploring Context-aware and LLM-driven Locomotion for Immersive Virtual Reality
by: Özdel, Süleyman, et al.
Published: (2025)
by: Özdel, Süleyman, et al.
Published: (2025)
From Passive Watching to Active Learning: Empowering Proactive Participation in Digital Classrooms with AI Video Assistant
by: Bodonhelyi, Anna, et al.
Published: (2024)
by: Bodonhelyi, Anna, et al.
Published: (2024)
TEyeD: Over 20 million real-world eye images with Pupil, Eyelid, and Iris 2D and 3D Segmentations, 2D and 3D Landmarks, 3D Eyeball, Gaze Vector, and Eye Movement Types
by: Fuhl, Wolfgang, et al.
Published: (2021)
by: Fuhl, Wolfgang, et al.
Published: (2021)
Understanding Password Preferences, Memorability, and Security through a Human-Centered Lens
by: Paker, Duru, et al.
Published: (2026)
by: Paker, Duru, et al.
Published: (2026)
Examining the Role of LLM-Driven Interactions on Attention and Cognitive Engagement in Virtual Classrooms
by: Ozdel, Suleyman, et al.
Published: (2025)
by: Ozdel, Suleyman, et al.
Published: (2025)
Emergent Abilities in Large Language Models: A Survey
by: Berti, Leonardo, et al.
Published: (2025)
by: Berti, Leonardo, et al.
Published: (2025)
Sure About That Line? Approaching Confidence-Based, Real-Time Line Assignment in Reading Gaze Data
by: Kaltenberger, Franziska, et al.
Published: (2026)
by: Kaltenberger, Franziska, et al.
Published: (2026)
Zero-Shot Pupil Segmentation with SAM 2: A Case Study of Over 14 Million Images
by: Maquiling, Virmarie, et al.
Published: (2024)
by: Maquiling, Virmarie, et al.
Published: (2024)
Mapping the Landscape of Artificial Intelligence in Life Cycle Assessment Using Large Language Models
by: Mensikova, Anastasija, et al.
Published: (2026)
by: Mensikova, Anastasija, et al.
Published: (2026)
Where Paths Split: Localized, Calibrated Control of Moral Reasoning in Large Language Models
by: Yuan, Chenchen, et al.
Published: (2026)
by: Yuan, Chenchen, et al.
Published: (2026)
Trade-offs in Privacy-Preserving Eye Tracking through Iris Obfuscation: A Benchmarking Study
by: Wang, Mengdi, et al.
Published: (2025)
by: Wang, Mengdi, et al.
Published: (2025)
What Shapes Participant Data Quality? A Scoping Review and Case Study of Crowdsourced Webcam Eye Tracking in AI Interviews
by: Lau, Ka Hei Carrie, et al.
Published: (2026)
by: Lau, Ka Hei Carrie, et al.
Published: (2026)
Imperfect Language, Artificial Intelligence, and the Human Mind: An Interdisciplinary Approach to Linguistic Errors in Native Spanish Speakers
by: López, Francisco Portillo
Published: (2025)
by: López, Francisco Portillo
Published: (2025)
Similar Items
-
Towards Adaptive Feedback with AI: Comparing the Feedback Quality of LLMs and Teachers on Experimentation Protocols
by: Seßler, Kathrin, et al.
Published: (2025) -
Taking the Next Step with Generative Artificial Intelligence: The Transformative Role of Multimodal Large Language Models in Science Education
by: Bewersdorff, Arne, et al.
Published: (2024) -
Benchmarking Large Language Models for Math Reasoning Tasks
by: Seßler, Kathrin, et al.
Published: (2024) -
Assessing Professional Development on Experimentation as a Method of Inquiry‐Based Science Teaching Framework and Principal Results
by: Markus Emden, et al.
Published: (2025) -
Can AI grade your essays? A comparative analysis of large language models and teacher ratings in multidimensional essay scoring
by: Seßler, Kathrin, et al.
Published: (2024)