SMART: Simulated Students Aligned with Item Response Theory for Question Difficulty Prediction
Fuente:
arXiv
Salvato in:
| Autori principali: | Scarlatos, Alexander, Fernandez, Nigel, Ormerod, Christopher, Lottridge, Susan, Lan, Andrew |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SyllabusQA: A Course Logistics Question Answering Dataset
di: Fernandez, Nigel, et al.
Pubblicazione: (2024)
di: Fernandez, Nigel, et al.
Pubblicazione: (2024)
DiVERT: Distractor Generation with Variational Errors Represented as Text for Math Multiple-choice Questions
di: Fernandez, Nigel, et al.
Pubblicazione: (2024)
di: Fernandez, Nigel, et al.
Pubblicazione: (2024)
Exploring Knowledge Tracing in Tutor-Student Dialogues using LLMs
di: Scarlatos, Alexander, et al.
Pubblicazione: (2024)
di: Scarlatos, Alexander, et al.
Pubblicazione: (2024)
KASER: Knowledge-Aligned Student Error Simulator for Open-Ended Coding Tasks
di: Duan, Zhangqi, et al.
Pubblicazione: (2026)
di: Duan, Zhangqi, et al.
Pubblicazione: (2026)
Interpreting Latent Student Knowledge Representations in Programming Assignments
di: Fernandez, Nigel, et al.
Pubblicazione: (2024)
di: Fernandez, Nigel, et al.
Pubblicazione: (2024)
Evaluating GPT-4 at Grading Handwritten Solutions in Math Exams
di: Caraeni, Adriana, et al.
Pubblicazione: (2024)
di: Caraeni, Adriana, et al.
Pubblicazione: (2024)
Exploring LLMs for Predicting Tutor Strategy and Student Outcomes in Dialogues
di: Ikram, Fareya, et al.
Pubblicazione: (2025)
di: Ikram, Fareya, et al.
Pubblicazione: (2025)
Simulated Students in Tutoring Dialogues: Substance or Illusion?
di: Scarlatos, Alexander, et al.
Pubblicazione: (2026)
di: Scarlatos, Alexander, et al.
Pubblicazione: (2026)
Test Case-Informed Knowledge Tracing for Open-ended Coding Tasks
di: Duan, Zhangqi, et al.
Pubblicazione: (2024)
di: Duan, Zhangqi, et al.
Pubblicazione: (2024)
Improving Automated Distractor Generation for Math Multiple-choice Questions with Overgenerate-and-rank
di: Scarlatos, Alexander, et al.
Pubblicazione: (2024)
di: Scarlatos, Alexander, et al.
Pubblicazione: (2024)
RetICL: Sequential Retrieval of In-Context Examples with Reinforcement Learning
di: Scarlatos, Alexander, et al.
Pubblicazione: (2023)
di: Scarlatos, Alexander, et al.
Pubblicazione: (2023)
Who Am I? History-Aware Profiles for Student Simulation in Tutoring Dialogues
di: Duan, Zhangqi, et al.
Pubblicazione: (2026)
di: Duan, Zhangqi, et al.
Pubblicazione: (2026)
Estimating Item Difficulty Using Large Language Models and Tree-Based Machine Learning Algorithms
di: Razavi, Pooya, et al.
Pubblicazione: (2025)
di: Razavi, Pooya, et al.
Pubblicazione: (2025)
Improving Socratic Question Generation using Data Augmentation and Preference Optimization
di: Kumar, Nischal Ashok, et al.
Pubblicazione: (2024)
di: Kumar, Nischal Ashok, et al.
Pubblicazione: (2024)
Training LLM-based Tutors to Improve Student Learning Outcomes in Dialogues
di: Scarlatos, Alexander, et al.
Pubblicazione: (2025)
di: Scarlatos, Alexander, et al.
Pubblicazione: (2025)
The Impact of Item-Writing Flaws on Difficulty and Discrimination in Item Response Theory
di: Schmucker, Robin, et al.
Pubblicazione: (2025)
di: Schmucker, Robin, et al.
Pubblicazione: (2025)
QG-SMS: Enhancing Test Item Analysis via Student Modeling and Simulation
di: Nguyen, Bang, et al.
Pubblicazione: (2025)
di: Nguyen, Bang, et al.
Pubblicazione: (2025)
Can LLMs Estimate Student Struggles? Human-AI Difficulty Alignment with Proficiency Simulation for Item Difficulty Prediction
di: Li, Ming, et al.
Pubblicazione: (2025)
di: Li, Ming, et al.
Pubblicazione: (2025)
UnibucLLM: Harnessing LLMs for Automated Prediction of Item Difficulty and Response Time for Multiple-Choice Questions
di: Rogoz, Ana-Cristina, et al.
Pubblicazione: (2024)
di: Rogoz, Ana-Cristina, et al.
Pubblicazione: (2024)
New Exam Security Questions in the AI Era: Comparing AI-Generated Item Similarity Between Naive and Detail-Guided Prompting Approaches
di: Wang, Ting, et al.
Pubblicazione: (2025)
di: Wang, Ting, et al.
Pubblicazione: (2025)
Interpretable Difficulty-Aware Knowledge Tracing in Tutor-Student Dialogues
di: Huang, Shuyan, et al.
Pubblicazione: (2026)
di: Huang, Shuyan, et al.
Pubblicazione: (2026)
Fairness Evaluation with Item Response Theory
di: Xu, Ziqi, et al.
Pubblicazione: (2024)
di: Xu, Ziqi, et al.
Pubblicazione: (2024)
Fairer Preferences Elicit Improved Human-Aligned Large Language Model Judgments
di: Zhou, Han, et al.
Pubblicazione: (2024)
di: Zhou, Han, et al.
Pubblicazione: (2024)
Embracing Imperfection: Simulating Students with Diverse Cognitive Levels Using LLM-based Agents
di: Wu, Tao, et al.
Pubblicazione: (2025)
di: Wu, Tao, et al.
Pubblicazione: (2025)
Student Answer Forecasting: Transformer-Driven Answer Choice Prediction for Language Learning
di: Gado, Elena Grazia, et al.
Pubblicazione: (2024)
di: Gado, Elena Grazia, et al.
Pubblicazione: (2024)
Automated Text Scoring in the Age of Generative AI for the GPU-poor
di: Ormerod, Christopher Michael, et al.
Pubblicazione: (2024)
di: Ormerod, Christopher Michael, et al.
Pubblicazione: (2024)
SimBench: Benchmarking the Ability of Large Language Models to Simulate Human Behaviors
di: Hu, Tiancheng, et al.
Pubblicazione: (2025)
di: Hu, Tiancheng, et al.
Pubblicazione: (2025)
Questionable practices in machine learning
di: Leech, Gavin, et al.
Pubblicazione: (2024)
di: Leech, Gavin, et al.
Pubblicazione: (2024)
TRIDENT: Benchmarking LLM Safety in Finance, Medicine, and Law
di: Hui, Zheng, et al.
Pubblicazione: (2025)
di: Hui, Zheng, et al.
Pubblicazione: (2025)
Reflecting in the Reflection: Integrating a Socratic Questioning Framework into Automated AI-Based Question Generation
di: Holub, Ondřej, et al.
Pubblicazione: (2026)
di: Holub, Ondřej, et al.
Pubblicazione: (2026)
Asking For It: Question-Answering for Predicting Rule Infractions in Online Content Moderation
di: Samory, Mattia, et al.
Pubblicazione: (2025)
di: Samory, Mattia, et al.
Pubblicazione: (2025)
Simulating Students or Sycophantic Problem Solving? On Misconception Faithfulness of LLM Simulators
di: Do, Heejin, et al.
Pubblicazione: (2026)
di: Do, Heejin, et al.
Pubblicazione: (2026)
ViMGuard: A Novel Multi-Modal System for Video Misinformation Guarding
di: Kan, Andrew, et al.
Pubblicazione: (2024)
di: Kan, Andrew, et al.
Pubblicazione: (2024)
Controlling Cloze-test Question Item Difficulty with PLM-based Surrogate Models for IRT Assessment
di: Zhang, Jingshen, et al.
Pubblicazione: (2024)
di: Zhang, Jingshen, et al.
Pubblicazione: (2024)
GG-BBQ: German Gender Bias Benchmark for Question Answering
di: Satheesh, Shalaka, et al.
Pubblicazione: (2025)
di: Satheesh, Shalaka, et al.
Pubblicazione: (2025)
Active Learning to Guide Labeling Efforts for Question Difficulty Estimation
di: Thuy, Arthur, et al.
Pubblicazione: (2024)
di: Thuy, Arthur, et al.
Pubblicazione: (2024)
Towards Unsupervised Question Answering System with Multi-level Summarization for Legal Text
di: Prabhu, M Manvith, et al.
Pubblicazione: (2024)
di: Prabhu, M Manvith, et al.
Pubblicazione: (2024)
Automated Knowledge Component Generation for Interpretable Knowledge Tracing in Coding Problems
di: Duan, Zhangqi, et al.
Pubblicazione: (2025)
di: Duan, Zhangqi, et al.
Pubblicazione: (2025)
Embedding Enhancement via Fine-Tuned Language Models for Learner-Item Cognitive Modeling
di: Liu, Yuanhao, et al.
Pubblicazione: (2026)
di: Liu, Yuanhao, et al.
Pubblicazione: (2026)
Representation Surgery: Theory and Practice of Affine Steering
di: Singh, Shashwat, et al.
Pubblicazione: (2024)
di: Singh, Shashwat, et al.
Pubblicazione: (2024)
Documenti analoghi
-
SyllabusQA: A Course Logistics Question Answering Dataset
di: Fernandez, Nigel, et al.
Pubblicazione: (2024) -
DiVERT: Distractor Generation with Variational Errors Represented as Text for Math Multiple-choice Questions
di: Fernandez, Nigel, et al.
Pubblicazione: (2024) -
Exploring Knowledge Tracing in Tutor-Student Dialogues using LLMs
di: Scarlatos, Alexander, et al.
Pubblicazione: (2024) -
KASER: Knowledge-Aligned Student Error Simulator for Open-Ended Coding Tasks
di: Duan, Zhangqi, et al.
Pubblicazione: (2026) -
Interpreting Latent Student Knowledge Representations in Programming Assignments
di: Fernandez, Nigel, et al.
Pubblicazione: (2024)