Enhancing Essay Cohesion Assessment: A Novel Item Response Theory Approach
Fuente:
arXiv
Saved in:
| Main Authors: | Rosa, Bruno Alexandre, Oliveira, Hilário, Rodrigues, Luiz, Oliveira, Eduardo Araujo, Mello, Rafael Ferreira |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Impact of Item-Writing Flaws on Difficulty and Discrimination in Item Response Theory
by: Schmucker, Robin, et al.
Published: (2025)
by: Schmucker, Robin, et al.
Published: (2025)
MAQuA: Adaptive Question-Asking for Multidimensional Mental Health Screening using Item Response Theory
by: Varadarajan, Vasudha, et al.
Published: (2025)
by: Varadarajan, Vasudha, et al.
Published: (2025)
JE-IRT: A Geometric Lens on LLM Abilities through Joint Embedding Item Response Theory
by: Yao, Louie Hong, et al.
Published: (2025)
by: Yao, Louie Hong, et al.
Published: (2025)
Rationale Behind Essay Scores: Enhancing S-LLM's Multi-Trait Essay Scoring with Rationale Generated by LLMs
by: Chu, SeongYeub, et al.
Published: (2024)
by: Chu, SeongYeub, et al.
Published: (2024)
Psychometric Item Validation Using Virtual Respondents with Trait-Response Mediators
by: Lim, Sungjib, et al.
Published: (2025)
by: Lim, Sungjib, et al.
Published: (2025)
Rank-Then-Score: Enhancing Large Language Models for Automated Essay Scoring
by: Cai, Yida, et al.
Published: (2025)
by: Cai, Yida, et al.
Published: (2025)
Hey AI Can You Grade My Essay?: Automatic Essay Grading
by: Maliha, Maisha, et al.
Published: (2024)
by: Maliha, Maisha, et al.
Published: (2024)
Enhancing Essay Scoring with Adversarial Weights Perturbation and Metric-specific AttentionPooling
by: Huang, Jiaxin, et al.
Published: (2024)
by: Huang, Jiaxin, et al.
Published: (2024)
From Test-Taking to Test-Making: Examining LLM Authoring of Commonsense Assessment Items
by: Roemmele, Melissa, et al.
Published: (2024)
by: Roemmele, Melissa, et al.
Published: (2024)
EssayBench: Evaluating Large Language Models in Multi-Genre Chinese Essay Writing
by: Gao, Fan, et al.
Published: (2025)
by: Gao, Fan, et al.
Published: (2025)
Evaluating Austrian A-Level German Essays with Large Language Models for Automated Essay Scoring
by: Kubesch, Jonas, et al.
Published: (2026)
by: Kubesch, Jonas, et al.
Published: (2026)
Large Language Models for Medical OSCE Assessment: A Novel Approach to Transcript Analysis
by: Shakur, Ameer Hamza, et al.
Published: (2024)
by: Shakur, Ameer Hamza, et al.
Published: (2024)
SLIM-RAFT: A Novel Fine-Tuning Approach to Improve Cross-Linguistic Performance for Mercosur Common Nomenclature
by: Di Oliveira, Vinícius, et al.
Published: (2024)
by: Di Oliveira, Vinícius, et al.
Published: (2024)
Human-AI Collaboration or Academic Misconduct? Measuring AI Use in Student Writing Through Stylometric Evidence
by: Oliveira, Eduardo Araujo, et al.
Published: (2025)
by: Oliveira, Eduardo Araujo, et al.
Published: (2025)
EssayJudge: A Multi-Granular Benchmark for Assessing Automated Essay Scoring Capabilities of Multimodal Large Language Models
by: Su, Jiamin, et al.
Published: (2025)
by: Su, Jiamin, et al.
Published: (2025)
Controlling Cloze-test Question Item Difficulty with PLM-based Surrogate Models for IRT Assessment
by: Zhang, Jingshen, et al.
Published: (2024)
by: Zhang, Jingshen, et al.
Published: (2024)
LLMs Do Not Grade Essays Like Humans
by: Mathew, Jerin George, et al.
Published: (2026)
by: Mathew, Jerin George, et al.
Published: (2026)
Enhancing Text Authenticity: A Novel Hybrid Approach for AI-Generated Text Detection
by: Zhang, Ye, et al.
Published: (2024)
by: Zhang, Ye, et al.
Published: (2024)
How well can LLMs Grade Essays in Arabic?
by: Ghazawi, Rayed, et al.
Published: (2025)
by: Ghazawi, Rayed, et al.
Published: (2025)
Autoregressive Score Generation for Multi-trait Essay Scoring
by: Do, Heejin, et al.
Published: (2024)
by: Do, Heejin, et al.
Published: (2024)
Automatic Reflection Level Classification in Hungarian Student Essays
by: Csibi, Zsolt, et al.
Published: (2026)
by: Csibi, Zsolt, et al.
Published: (2026)
Tagarela - A Portuguese speech dataset from podcasts
by: de Oliveira, Frederico Santos, et al.
Published: (2026)
by: de Oliveira, Frederico Santos, et al.
Published: (2026)
`Keep it Together': Enforcing Cohesion in Extractive Summaries by Simulating Human Memory
by: Cardenas, Ronald, et al.
Published: (2024)
by: Cardenas, Ronald, et al.
Published: (2024)
DREsS: Dataset for Rubric-based Essay Scoring on EFL Writing
by: Yoo, Haneul, et al.
Published: (2024)
by: Yoo, Haneul, et al.
Published: (2024)
Activations as Features: Probing LLMs for Generalizable Essay Scoring Representations
by: Chi, Jinwei, et al.
Published: (2025)
by: Chi, Jinwei, et al.
Published: (2025)
Improve LLM-based Automatic Essay Scoring with Linguistic Features
by: Hou, Zhaoyi Joey, et al.
Published: (2025)
by: Hou, Zhaoyi Joey, et al.
Published: (2025)
Automatic Essay Scoring and Feedback Generation in Basque Language Learning
by: Azurmendi, Ekhi, et al.
Published: (2025)
by: Azurmendi, Ekhi, et al.
Published: (2025)
Are Large Language Models Good Essay Graders?
by: Kundu, Anindita, et al.
Published: (2024)
by: Kundu, Anindita, et al.
Published: (2024)
Teaching-Inspired Integrated Prompting Framework: A Novel Approach for Enhancing Reasoning in Large Language Models
by: Tan, Wenting, et al.
Published: (2024)
by: Tan, Wenting, et al.
Published: (2024)
Assessing GPTZero's Accuracy in Identifying AI vs. Human-Written Essays
by: Dik, Selin, et al.
Published: (2025)
by: Dik, Selin, et al.
Published: (2025)
Automated Essay Scoring Incorporating Annotations from Automated Feedback Systems
by: Ormerod, Christopher
Published: (2025)
by: Ormerod, Christopher
Published: (2025)
Hidding the Ghostwriters: An Adversarial Evaluation of AI-Generated Student Essay Detection
by: Peng, Xinlin, et al.
Published: (2024)
by: Peng, Xinlin, et al.
Published: (2024)
Automatic Essay Multi-dimensional Scoring with Fine-tuning and Multiple Regression
by: Sun, Kun, et al.
Published: (2024)
by: Sun, Kun, et al.
Published: (2024)
Can Large Language Models Automatically Score Proficiency of Written Essays?
by: Mansour, Watheq, et al.
Published: (2024)
by: Mansour, Watheq, et al.
Published: (2024)
AI-generated Essays: Characteristics and Implications on Automated Scoring and Academic Integrity
by: Zhong, Yang, et al.
Published: (2024)
by: Zhong, Yang, et al.
Published: (2024)
Theory of Mind in Large Language Models: Assessment and Enhancement
by: Chen, Ruirui, et al.
Published: (2025)
by: Chen, Ruirui, et al.
Published: (2025)
Modifying AI, Enhancing Essays: How Active Engagement with Generative AI Boosts Writing Quality
by: Yang, Kaixun, et al.
Published: (2024)
by: Yang, Kaixun, et al.
Published: (2024)
Using ChatGPT to Score Essays and Short-Form Constructed Responses
by: Shermis, Mark D.
Published: (2024)
by: Shermis, Mark D.
Published: (2024)
A Novel Computational and Modeling Foundation for Automatic Coherence Assessment
by: Maimon, Aviya, et al.
Published: (2023)
by: Maimon, Aviya, et al.
Published: (2023)
Prove Your Point!: Bringing Proof-Enhancement Principles to Argumentative Essay Generation
by: Xiao, Ruiyu, et al.
Published: (2024)
by: Xiao, Ruiyu, et al.
Published: (2024)
Similar Items
-
The Impact of Item-Writing Flaws on Difficulty and Discrimination in Item Response Theory
by: Schmucker, Robin, et al.
Published: (2025) -
MAQuA: Adaptive Question-Asking for Multidimensional Mental Health Screening using Item Response Theory
by: Varadarajan, Vasudha, et al.
Published: (2025) -
JE-IRT: A Geometric Lens on LLM Abilities through Joint Embedding Item Response Theory
by: Yao, Louie Hong, et al.
Published: (2025) -
Rationale Behind Essay Scores: Enhancing S-LLM's Multi-Trait Essay Scoring with Rationale Generated by LLMs
by: Chu, SeongYeub, et al.
Published: (2024) -
Psychometric Item Validation Using Virtual Respondents with Trait-Response Mediators
by: Lim, Sungjib, et al.
Published: (2025)