Leveraging LLM-Respondents for Item Evaluation: a Psychometric Analysis
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Yunting, Bhandari, Shreya, Pardos, Zachary A. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Psychometric Item Validation Using Virtual Respondents with Trait-Response Mediators
by: Lim, Sungjib, et al.
Published: (2025)
by: Lim, Sungjib, et al.
Published: (2025)
The RIGID Framework: Research-Integrated, Generative AI-Mediated Instructional Design
by: Kwak, Yerin, et al.
Published: (2026)
by: Kwak, Yerin, et al.
Published: (2026)
The Impact of Item-Writing Flaws on Difficulty and Discrimination in Item Response Theory
by: Schmucker, Robin, et al.
Published: (2025)
by: Schmucker, Robin, et al.
Published: (2025)
Psychometric Comparability of LLM-Based Digital Twins
by: Zhang, Yufei, et al.
Published: (2025)
by: Zhang, Yufei, et al.
Published: (2025)
NLP Cluster Analysis of Common Core State Standards and NAEP Item Specifications
by: Camilli, Gregory, et al.
Published: (2024)
by: Camilli, Gregory, et al.
Published: (2024)
PromptHive: Bringing Subject Matter Experts Back to the Forefront with Collaborative Prompt Engineering for Educational Content Creation
by: Reza, Mohi, et al.
Published: (2024)
by: Reza, Mohi, et al.
Published: (2024)
Leveraging LLM respondents for item evaluation: A psychometric analysis
by: Yunting Liu, et al.
Published: (2025)
by: Yunting Liu, et al.
Published: (2025)
Advancing credit mobility through stakeholder-informed AI design and adoption
by: Kwak, Yerin, et al.
Published: (2026)
by: Kwak, Yerin, et al.
Published: (2026)
Teaching at Scale: Leveraging AI to Evaluate and Elevate Engineering Education
by: Chamberland, Jean-Francois, et al.
Published: (2025)
by: Chamberland, Jean-Francois, et al.
Published: (2025)
SHIELD: Evaluation and Defense Strategies for Copyright Compliance in LLM Text Generation
by: Liu, Xiaoze, et al.
Published: (2024)
by: Liu, Xiaoze, et al.
Published: (2024)
Evaluating how LLM annotations represent diverse views on contentious topics
by: Brown, Megan A., et al.
Published: (2025)
by: Brown, Megan A., et al.
Published: (2025)
QG-SMS: Enhancing Test Item Analysis via Student Modeling and Simulation
by: Nguyen, Bang, et al.
Published: (2025)
by: Nguyen, Bang, et al.
Published: (2025)
Political Alignment in Large Language Models: A Multidimensional Audit of Psychometric Identity and Behavioral Bias
by: Sakhawat, Adib, et al.
Published: (2026)
by: Sakhawat, Adib, et al.
Published: (2026)
Controlling Cloze-test Question Item Difficulty with PLM-based Surrogate Models for IRT Assessment
by: Zhang, Jingshen, et al.
Published: (2024)
by: Zhang, Jingshen, et al.
Published: (2024)
Do GPT Language Models Suffer From Split Personality Disorder? The Advent Of Substrate-Free Psychometrics
by: Romero, Peter, et al.
Published: (2024)
by: Romero, Peter, et al.
Published: (2024)
Augmenting Rating-Scale Measures with Text-Derived Items Using the Information-Determined Scoring (IDS) Framework
by: Watson, Joe, et al.
Published: (2025)
by: Watson, Joe, et al.
Published: (2025)
Evaluation of LLM Vulnerabilities to Being Misused for Personalized Disinformation Generation
by: Zugecova, Aneta, et al.
Published: (2024)
by: Zugecova, Aneta, et al.
Published: (2024)
In-Situ Behavioral Evaluation for LLM Fairness, Not Standardized-Test Scores
by: Tang, Zeyu, et al.
Published: (2026)
by: Tang, Zeyu, et al.
Published: (2026)
Sacred or Synthetic? Evaluating LLM Reliability and Abstention for Religious Questions
by: Atif, Farah, et al.
Published: (2025)
by: Atif, Farah, et al.
Published: (2025)
Can LLMs Estimate Student Struggles? Human-AI Difficulty Alignment with Proficiency Simulation for Item Difficulty Prediction
by: Li, Ming, et al.
Published: (2025)
by: Li, Ming, et al.
Published: (2025)
Machine Learning for Detection and Analysis of Novel LLM Jailbreaks
by: Hawkins, John, et al.
Published: (2025)
by: Hawkins, John, et al.
Published: (2025)
Evaluating Implicit Biases in LLM Reasoning through Logic Grid Puzzles
by: Jahara, Fatima, et al.
Published: (2025)
by: Jahara, Fatima, et al.
Published: (2025)
LLM BiasScope: A Real-Time Bias Analysis Platform for Comparative LLM Evaluation
by: Ghosh, Himel, et al.
Published: (2026)
by: Ghosh, Himel, et al.
Published: (2026)
Automated Item Neutralization for Non-Cognitive Scales: A Large Language Model Approach to Reducing Social-Desirability Bias
by: Wu, Sirui, et al.
Published: (2025)
by: Wu, Sirui, et al.
Published: (2025)
From Demographics to Survey Anchors: Evaluating LLM Agents for Modeling Retirement Attitudes
by: Garzón, Rubén, et al.
Published: (2026)
by: Garzón, Rubén, et al.
Published: (2026)
Embedding Enhancement via Fine-Tuned Language Models for Learner-Item Cognitive Modeling
by: Liu, Yuanhao, et al.
Published: (2026)
by: Liu, Yuanhao, et al.
Published: (2026)
Mass-Scale Analysis of In-the-Wild Conversations Reveals Complexity Bounds on LLM Jailbreaking
by: Creo, Aldan, et al.
Published: (2025)
by: Creo, Aldan, et al.
Published: (2025)
Detecting AI-Generated Text in Educational Content: Leveraging Machine Learning and Explainable AI for Academic Integrity
by: Najjar, Ayat A., et al.
Published: (2025)
by: Najjar, Ayat A., et al.
Published: (2025)
Gender and Positional Biases in LLM-Based Hiring Decisions: Evidence from Comparative CV/Résumé Evaluations
by: Rozado, David
Published: (2025)
by: Rozado, David
Published: (2025)
Bias in the Tails: How Name-conditioned Evaluative Framing in Resume Summaries Destabilizes LLM-based Hiring
by: Nghiem, Huy, et al.
Published: (2026)
by: Nghiem, Huy, et al.
Published: (2026)
Leveraging Prompts in LLMs to Overcome Imbalances in Complex Educational Text Data
by: McClure, Jeanne, et al.
Published: (2024)
by: McClure, Jeanne, et al.
Published: (2024)
Evaluating the Impact of Advanced LLM Techniques on AI-Lecture Tutors for a Robotics Course
by: Kahl, Sebastian, et al.
Published: (2024)
by: Kahl, Sebastian, et al.
Published: (2024)
Think Like a Person Before Responding: A Multi-Faceted Evaluation of Persona-Guided LLMs for Countering Hate
by: Ngueajio, Mikel K., et al.
Published: (2025)
by: Ngueajio, Mikel K., et al.
Published: (2025)
PersLLM: A Personified Training Approach for Large Language Models
by: Zeng, Zheni, et al.
Published: (2024)
by: Zeng, Zheni, et al.
Published: (2024)
Leveraging Lecture Content for Improved Feedback: Explorations with GPT-4 and Retrieval Augmented Generation
by: Jacobs, Sven, et al.
Published: (2024)
by: Jacobs, Sven, et al.
Published: (2024)
Leveraging Social Determinants of Health in Alzheimer's Research Using LLM-Augmented Literature Mining and Knowledge Graphs
by: Shang, Tianqi, et al.
Published: (2024)
by: Shang, Tianqi, et al.
Published: (2024)
Large Language Models Leverage External Knowledge to Extend Clinical Insight Beyond Language Boundaries
by: Wu, Jiageng, et al.
Published: (2023)
by: Wu, Jiageng, et al.
Published: (2023)
A Scoping Review of LLM-as-a-Judge in Healthcare and the MedJUDGE Framework
by: Li, Chenyu, et al.
Published: (2026)
by: Li, Chenyu, et al.
Published: (2026)
Mental Health Equity in LLMs: Leveraging Multi-Hop Question Answering to Detect Amplified and Silenced Perspectives
by: Haider, Batool, et al.
Published: (2025)
by: Haider, Batool, et al.
Published: (2025)
Intelligent Tutor: Leveraging ChatGPT and Microsoft Copilot Studio to Deliver a Generative AI Student Support and Feedback System within Teams
by: Chen, Wei-Yu
Published: (2024)
by: Chen, Wei-Yu
Published: (2024)
Similar Items
-
Psychometric Item Validation Using Virtual Respondents with Trait-Response Mediators
by: Lim, Sungjib, et al.
Published: (2025) -
The RIGID Framework: Research-Integrated, Generative AI-Mediated Instructional Design
by: Kwak, Yerin, et al.
Published: (2026) -
The Impact of Item-Writing Flaws on Difficulty and Discrimination in Item Response Theory
by: Schmucker, Robin, et al.
Published: (2025) -
Psychometric Comparability of LLM-Based Digital Twins
by: Zhang, Yufei, et al.
Published: (2025) -
NLP Cluster Analysis of Common Core State Standards and NAEP Item Specifications
by: Camilli, Gregory, et al.
Published: (2024)