Beyond Agreement: Diagnosing the Rationale Alignment of Automated Essay Scoring Methods based on Linguistically-informed Counterfactuals
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Yupei, Hu, Renfen, Zhao, Zhe |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Does Visual Grounding Enhance the Understanding of Embodied Knowledge in Large Language Models?
von: Yang, Zhihui, et al.
Veröffentlicht: (2025)
von: Yang, Zhihui, et al.
Veröffentlicht: (2025)
Rationale Behind Essay Scores: Enhancing S-LLM's Multi-Trait Essay Scoring with Rationale Generated by LLMs
von: Chu, SeongYeub, et al.
Veröffentlicht: (2024)
von: Chu, SeongYeub, et al.
Veröffentlicht: (2024)
Automated Essay Scoring and Language Certification: Assessing Generalizability, Agreement and Validity for French
von: Wilkens, Rodrigo, et al.
Veröffentlicht: (2026)
von: Wilkens, Rodrigo, et al.
Veröffentlicht: (2026)
Beyond the Score: Uncertainty-Calibrated LLMs for Automated Essay Assessment
von: Karim, Ahmed, et al.
Veröffentlicht: (2025)
von: Karim, Ahmed, et al.
Veröffentlicht: (2025)
Improve LLM-based Automatic Essay Scoring with Linguistic Features
von: Hou, Zhaoyi Joey, et al.
Veröffentlicht: (2025)
von: Hou, Zhaoyi Joey, et al.
Veröffentlicht: (2025)
Deep Associations, High Creativity: A Simple yet Effective Metric for Evaluating Large Language Models
von: Qiu, Ziliang, et al.
Veröffentlicht: (2025)
von: Qiu, Ziliang, et al.
Veröffentlicht: (2025)
Teach-to-Reason with Scoring: Self-Explainable Rationale-Driven Multi-Trait Essay Scoring
von: Do, Heejin, et al.
Veröffentlicht: (2025)
von: Do, Heejin, et al.
Veröffentlicht: (2025)
Efficiently Building a Domain-Specific Large Language Model from Scratch: A Case Study of a Classical Chinese Large Language Model
von: Li, Shen, et al.
Veröffentlicht: (2025)
von: Li, Shen, et al.
Veröffentlicht: (2025)
Rank-Then-Score: Enhancing Large Language Models for Automated Essay Scoring
von: Cai, Yida, et al.
Veröffentlicht: (2025)
von: Cai, Yida, et al.
Veröffentlicht: (2025)
Long Context Automated Essay Scoring with Language Models
von: Ormerod, Christopher, et al.
Veröffentlicht: (2025)
von: Ormerod, Christopher, et al.
Veröffentlicht: (2025)
Beyond Holistic Scores: Automatic Trait-Based Quality Scoring of Argumentative Essays
von: Favero, Lucile, et al.
Veröffentlicht: (2026)
von: Favero, Lucile, et al.
Veröffentlicht: (2026)
Adversarial Topic-aware Prompt-tuning for Cross-topic Automated Essay Scoring
von: Zhang, Chunyun, et al.
Veröffentlicht: (2025)
von: Zhang, Chunyun, et al.
Veröffentlicht: (2025)
Automated Essay Scoring Incorporating Annotations from Automated Feedback Systems
von: Ormerod, Christopher
Veröffentlicht: (2025)
von: Ormerod, Christopher
Veröffentlicht: (2025)
Enhancing Automated Essay Scoring with Three Techniques: Two-Stage Fine-Tuning, Score Alignment, and Self-Training
von: Choi, Hongseok, et al.
Veröffentlicht: (2026)
von: Choi, Hongseok, et al.
Veröffentlicht: (2026)
Exploration of Summarization by Generative Language Models for Automated Scoring of Long Essays
von: Hua, Haowei, et al.
Veröffentlicht: (2025)
von: Hua, Haowei, et al.
Veröffentlicht: (2025)
Enhancing Arabic Automated Essay Scoring with Synthetic Data and Error Injection
von: Qwaider, Chatrine, et al.
Veröffentlicht: (2025)
von: Qwaider, Chatrine, et al.
Veröffentlicht: (2025)
AI-generated Essays: Characteristics and Implications on Automated Scoring and Academic Integrity
von: Zhong, Yang, et al.
Veröffentlicht: (2024)
von: Zhong, Yang, et al.
Veröffentlicht: (2024)
Evaluating Austrian A-Level German Essays with Large Language Models for Automated Essay Scoring
von: Kubesch, Jonas, et al.
Veröffentlicht: (2026)
von: Kubesch, Jonas, et al.
Veröffentlicht: (2026)
A Rationale-centric Counterfactual Data Augmentation Method for Cross-Document Event Coreference Resolution
von: Ding, Bowen, et al.
Veröffentlicht: (2024)
von: Ding, Bowen, et al.
Veröffentlicht: (2024)
EssayJudge: A Multi-Granular Benchmark for Assessing Automated Essay Scoring Capabilities of Multimodal Large Language Models
von: Su, Jiamin, et al.
Veröffentlicht: (2025)
von: Su, Jiamin, et al.
Veröffentlicht: (2025)
Operationalizing Automated Essay Scoring: A Human-Aware Approach
von: Plasencia-Calaña, Yenisel
Veröffentlicht: (2025)
von: Plasencia-Calaña, Yenisel
Veröffentlicht: (2025)
IELTS Writing Revision Platform with Automated Essay Scoring and Adaptive Feedback
von: Ramancauskas, Titas, et al.
Veröffentlicht: (2025)
von: Ramancauskas, Titas, et al.
Veröffentlicht: (2025)
Automated Refinement of Essay Scoring Rubrics for Language Models via Reflect-and-Revise
von: Harada, Keno, et al.
Veröffentlicht: (2025)
von: Harada, Keno, et al.
Veröffentlicht: (2025)
LAILA: A Large Trait-Based Dataset for Arabic Automated Essay Scoring
von: Bashendy, May, et al.
Veröffentlicht: (2025)
von: Bashendy, May, et al.
Veröffentlicht: (2025)
MARE: Multi-Aspect Rationale Extractor on Unsupervised Rationale Extraction
von: Jiang, Han, et al.
Veröffentlicht: (2024)
von: Jiang, Han, et al.
Veröffentlicht: (2024)
Unveiling the Tapestry of Automated Essay Scoring: A Comprehensive Investigation of Accuracy, Fairness, and Generalizability
von: Yang, Kaixun, et al.
Veröffentlicht: (2024)
von: Yang, Kaixun, et al.
Veröffentlicht: (2024)
Empirical Analysis of the Effect of Context in the Task of Automated Essay Scoring in Transformer-Based Models
von: Chakravarty, Abhirup
Veröffentlicht: (2025)
von: Chakravarty, Abhirup
Veröffentlicht: (2025)
Investigating the Impact of Rationales for LLMs on Natural Language Understanding
von: Shi, Wenhang, et al.
Veröffentlicht: (2025)
von: Shi, Wenhang, et al.
Veröffentlicht: (2025)
Towards Prompt Generalization: Grammar-aware Cross-Prompt Automated Essay Scoring
von: Do, Heejin, et al.
Veröffentlicht: (2025)
von: Do, Heejin, et al.
Veröffentlicht: (2025)
The Impact of Example Selection in Few-Shot Prompting on Automated Essay Scoring Using GPT Models
von: Yoshida, Lui
Veröffentlicht: (2024)
von: Yoshida, Lui
Veröffentlicht: (2024)
Beyond Agreement: Scoring Panel-Surfaced Biomedical Entity Candidates for Curator Triage
von: Cao, Shuheng, et al.
Veröffentlicht: (2026)
von: Cao, Shuheng, et al.
Veröffentlicht: (2026)
Automated Essay Scoring Using Grammatical Variety and Errors with Multi-Task Learning and Item Response Theory
von: Doi, Kosuke, et al.
Veröffentlicht: (2024)
von: Doi, Kosuke, et al.
Veröffentlicht: (2024)
Do We Need a Detailed Rubric for Automated Essay Scoring using Large Language Models?
von: Yoshida, Lui
Veröffentlicht: (2025)
von: Yoshida, Lui
Veröffentlicht: (2025)
Automatic Essay Scoring in a Brazilian Scenario
von: Matsuoka, Felipe Akio
Veröffentlicht: (2023)
von: Matsuoka, Felipe Akio
Veröffentlicht: (2023)
Autoregressive Score Generation for Multi-trait Essay Scoring
von: Do, Heejin, et al.
Veröffentlicht: (2024)
von: Do, Heejin, et al.
Veröffentlicht: (2024)
DREsS: Dataset for Rubric-based Essay Scoring on EFL Writing
von: Yoo, Haneul, et al.
Veröffentlicht: (2024)
von: Yoo, Haneul, et al.
Veröffentlicht: (2024)
Align-GRAG: Anchor and Rationale Guided Dual Alignment for Graph Retrieval-Augmented Generation
von: Xu, Derong, et al.
Veröffentlicht: (2025)
von: Xu, Derong, et al.
Veröffentlicht: (2025)
Bridging the LLM Accessibility Divide? Performance, Fairness, and Cost of Closed versus Open LLMs for Automated Essay Scoring
von: Oketch, Kezia, et al.
Veröffentlicht: (2025)
von: Oketch, Kezia, et al.
Veröffentlicht: (2025)
Trait-Aware Policy Optimization for Autoregressive Multi-Trait Essay Scoring
von: Wang, Zhengyang, et al.
Veröffentlicht: (2026)
von: Wang, Zhengyang, et al.
Veröffentlicht: (2026)
Is GPT-4 Alone Sufficient for Automated Essay Scoring?: A Comparative Judgment Approach Based on Rater Cognition
von: Kim, Seungju, et al.
Veröffentlicht: (2024)
von: Kim, Seungju, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Does Visual Grounding Enhance the Understanding of Embodied Knowledge in Large Language Models?
von: Yang, Zhihui, et al.
Veröffentlicht: (2025) -
Rationale Behind Essay Scores: Enhancing S-LLM's Multi-Trait Essay Scoring with Rationale Generated by LLMs
von: Chu, SeongYeub, et al.
Veröffentlicht: (2024) -
Automated Essay Scoring and Language Certification: Assessing Generalizability, Agreement and Validity for French
von: Wilkens, Rodrigo, et al.
Veröffentlicht: (2026) -
Beyond the Score: Uncertainty-Calibrated LLMs for Automated Essay Assessment
von: Karim, Ahmed, et al.
Veröffentlicht: (2025) -
Improve LLM-based Automatic Essay Scoring with Linguistic Features
von: Hou, Zhaoyi Joey, et al.
Veröffentlicht: (2025)