The Impact of Item-Writing Flaws on Difficulty and Discrimination in Item Response Theory
Fuente:
arXiv
Saved in:
| Main Authors: | Schmucker, Robin, Moore, Steven |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Can LLMs Estimate Student Struggles? Human-AI Difficulty Alignment with Proficiency Simulation for Item Difficulty Prediction
by: Li, Ming, et al.
Published: (2025)
by: Li, Ming, et al.
Published: (2025)
Controlling Cloze-test Question Item Difficulty with PLM-based Surrogate Models for IRT Assessment
by: Zhang, Jingshen, et al.
Published: (2024)
by: Zhang, Jingshen, et al.
Published: (2024)
Leveraging LLM-Respondents for Item Evaluation: a Psychometric Analysis
by: Liu, Yunting, et al.
Published: (2024)
by: Liu, Yunting, et al.
Published: (2024)
SMART: Simulated Students Aligned with Item Response Theory for Question Difficulty Prediction
by: Scarlatos, Alexander, et al.
Published: (2025)
by: Scarlatos, Alexander, et al.
Published: (2025)
NLP Cluster Analysis of Common Core State Standards and NAEP Item Specifications
by: Camilli, Gregory, et al.
Published: (2024)
by: Camilli, Gregory, et al.
Published: (2024)
Automated Generation and Tagging of Knowledge Components from Multiple-Choice Questions
by: Moore, Steven, et al.
Published: (2024)
by: Moore, Steven, et al.
Published: (2024)
Augmenting Rating-Scale Measures with Text-Derived Items Using the Information-Determined Scoring (IDS) Framework
by: Watson, Joe, et al.
Published: (2025)
by: Watson, Joe, et al.
Published: (2025)
Automated Item Neutralization for Non-Cognitive Scales: A Large Language Model Approach to Reducing Social-Desirability Bias
by: Wu, Sirui, et al.
Published: (2025)
by: Wu, Sirui, et al.
Published: (2025)
QG-SMS: Enhancing Test Item Analysis via Student Modeling and Simulation
by: Nguyen, Bang, et al.
Published: (2025)
by: Nguyen, Bang, et al.
Published: (2025)
Embedding Enhancement via Fine-Tuned Language Models for Learner-Item Cognitive Modeling
by: Liu, Yuanhao, et al.
Published: (2026)
by: Liu, Yuanhao, et al.
Published: (2026)
Enhancing Essay Cohesion Assessment: A Novel Item Response Theory Approach
by: Rosa, Bruno Alexandre, et al.
Published: (2025)
by: Rosa, Bruno Alexandre, et al.
Published: (2025)
UnibucLLM: Harnessing LLMs for Automated Prediction of Item Difficulty and Response Time for Multiple-Choice Questions
by: Rogoz, Ana-Cristina, et al.
Published: (2024)
by: Rogoz, Ana-Cristina, et al.
Published: (2024)
MAQuA: Adaptive Question-Asking for Multidimensional Mental Health Screening using Item Response Theory
by: Varadarajan, Vasudha, et al.
Published: (2025)
by: Varadarajan, Vasudha, et al.
Published: (2025)
JE-IRT: A Geometric Lens on LLM Abilities through Joint Embedding Item Response Theory
by: Yao, Louie Hong, et al.
Published: (2025)
by: Yao, Louie Hong, et al.
Published: (2025)
AI Mentors for Student Projects: Spotting Early Issues in Computer Science Proposals
by: Aher, Gati, et al.
Published: (2025)
by: Aher, Gati, et al.
Published: (2025)
Psychometric Item Validation Using Virtual Respondents with Trait-Response Mediators
by: Lim, Sungjib, et al.
Published: (2025)
by: Lim, Sungjib, et al.
Published: (2025)
Using Vision + Language Models to Predict Item Difficulty
by: Khan, Samin
Published: (2026)
by: Khan, Samin
Published: (2026)
QueerBench: Quantifying Discrimination in Language Models Toward Queer Identities
by: Sosto, Mae, et al.
Published: (2024)
by: Sosto, Mae, et al.
Published: (2024)
Estimating Item Difficulty Using Large Language Models and Tree-Based Machine Learning Algorithms
by: Razavi, Pooya, et al.
Published: (2025)
by: Razavi, Pooya, et al.
Published: (2025)
Writing in Symbiosis: Mapping Human Creative Agency in the AI Era
by: Doshi, Vivan, et al.
Published: (2025)
by: Doshi, Vivan, et al.
Published: (2025)
Optimizing Mastery Learning by Fast-Forwarding Over-Practice Steps
by: Xia, Meng, et al.
Published: (2025)
by: Xia, Meng, et al.
Published: (2025)
Authorship Without Writing: Large Language Models and the Senior Author Analogy
by: Hurshman, Clint, et al.
Published: (2025)
by: Hurshman, Clint, et al.
Published: (2025)
RIDE: Difficulty Evolving Perturbation with Item Response Theory for Mathematical Reasoning
by: Li, Xinyuan, et al.
Published: (2025)
by: Li, Xinyuan, et al.
Published: (2025)
ASCenD-BDS: Adaptable, Stochastic and Context-aware framework for Detection of Bias, Discrimination and Stereotyping
by: Bahl, Rajiv, et al.
Published: (2025)
by: Bahl, Rajiv, et al.
Published: (2025)
Prediction of Item Difficulty for Reading Comprehension Items by Creation of Annotated Item Repository
by: Kapoor, Radhika, et al.
Published: (2025)
by: Kapoor, Radhika, et al.
Published: (2025)
LLM-Driven Robots Risk Enacting Discrimination, Violence, and Unlawful Actions
by: Hundt, Andrew, et al.
Published: (2024)
by: Hundt, Andrew, et al.
Published: (2024)
Human-AI Collaboration or Academic Misconduct? Measuring AI Use in Student Writing Through Stylometric Evidence
by: Oliveira, Eduardo Araujo, et al.
Published: (2025)
by: Oliveira, Eduardo Araujo, et al.
Published: (2025)
AI Evaluation Should Require Standardized Item-Level Data Releases
by: Jiang, Han, et al.
Published: (2026)
by: Jiang, Han, et al.
Published: (2026)
The Course Difficulty Analysis Cookbook
by: Baucks, Frederik, et al.
Published: (2025)
by: Baucks, Frederik, et al.
Published: (2025)
The Impact of Large Language Models in Academia: from Writing to Speaking
by: Geng, Mingmeng, et al.
Published: (2024)
by: Geng, Mingmeng, et al.
Published: (2024)
Measuring Competency, Not Performance: Item-Aware Evaluation Across Medical Benchmarks
by: Luo, Zhimeng, et al.
Published: (2025)
by: Luo, Zhimeng, et al.
Published: (2025)
Confident Rankings with Fewer Items: Adaptive LLM Evaluation with Continuous Scores
by: Balkır, Esma, et al.
Published: (2026)
by: Balkır, Esma, et al.
Published: (2026)
Computational Approaches to Understanding Large Language Model Impact on Writing and Information Ecosystems
by: Liang, Weixin
Published: (2025)
by: Liang, Weixin
Published: (2025)
The Responsible Development of Automated Student Feedback with Generative AI
by: Lindsay, Euan D, et al.
Published: (2023)
by: Lindsay, Euan D, et al.
Published: (2023)
Conformity and Social Impact on AI Agents
by: Bellina, Alessandro, et al.
Published: (2026)
by: Bellina, Alessandro, et al.
Published: (2026)
MIRA: A Bilingual Benchmark for Medical Information Response Audit
by: Xu, Mengyu, et al.
Published: (2026)
by: Xu, Mengyu, et al.
Published: (2026)
Medical Hallucinations in Foundation Models and Their Impact on Healthcare
by: Kim, Yubin, et al.
Published: (2025)
by: Kim, Yubin, et al.
Published: (2025)
"I Am the One and Only, Your Cyber BFF": Understanding the Impact of GenAI Requires Understanding the Impact of Anthropomorphic AI
by: Cheng, Myra, et al.
Published: (2024)
by: Cheng, Myra, et al.
Published: (2024)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
by: Peters, Sydney, et al.
Published: (2025)
by: Peters, Sydney, et al.
Published: (2025)
An NLP Crosswalk Between the Common Core State Standards and NAEP Item Specifications
by: Camilli, Gregory
Published: (2024)
by: Camilli, Gregory
Published: (2024)
Similar Items
-
Can LLMs Estimate Student Struggles? Human-AI Difficulty Alignment with Proficiency Simulation for Item Difficulty Prediction
by: Li, Ming, et al.
Published: (2025) -
Controlling Cloze-test Question Item Difficulty with PLM-based Surrogate Models for IRT Assessment
by: Zhang, Jingshen, et al.
Published: (2024) -
Leveraging LLM-Respondents for Item Evaluation: a Psychometric Analysis
by: Liu, Yunting, et al.
Published: (2024) -
SMART: Simulated Students Aligned with Item Response Theory for Question Difficulty Prediction
by: Scarlatos, Alexander, et al.
Published: (2025) -
NLP Cluster Analysis of Common Core State Standards and NAEP Item Specifications
by: Camilli, Gregory, et al.
Published: (2024)