Operationalizing Automated Essay Scoring: A Human-Aware Approach
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Plasencia-Calaña, Yenisel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RDBE: Reasoning Distillation-Based Evaluation Enhances Automatic Essay Scoring
von: Mohammadkhani, Ali Ghiasvand
Veröffentlicht: (2024)
von: Mohammadkhani, Ali Ghiasvand
Veröffentlicht: (2024)
Beyond the Score: Uncertainty-Calibrated LLMs for Automated Essay Assessment
von: Karim, Ahmed, et al.
Veröffentlicht: (2025)
von: Karim, Ahmed, et al.
Veröffentlicht: (2025)
Exploration of Summarization by Generative Language Models for Automated Scoring of Long Essays
von: Hua, Haowei, et al.
Veröffentlicht: (2025)
von: Hua, Haowei, et al.
Veröffentlicht: (2025)
Is GPT-4 Alone Sufficient for Automated Essay Scoring?: A Comparative Judgment Approach Based on Rater Cognition
von: Kim, Seungju, et al.
Veröffentlicht: (2024)
von: Kim, Seungju, et al.
Veröffentlicht: (2024)
A Time-Aware Approach to Early Detection of Anorexia: UNSL at eRisk 2024
von: Thompson, Horacio, et al.
Veröffentlicht: (2024)
von: Thompson, Horacio, et al.
Veröffentlicht: (2024)
Empirical Comparison of Encoder-Based Language Models and Feature-Based Supervised Machine Learning Approaches to Automated Scoring of Long Essays
von: Wang, Kuo, et al.
Veröffentlicht: (2026)
von: Wang, Kuo, et al.
Veröffentlicht: (2026)
Empirical Analysis of the Effect of Context in the Task of Automated Essay Scoring in Transformer-Based Models
von: Chakravarty, Abhirup
Veröffentlicht: (2025)
von: Chakravarty, Abhirup
Veröffentlicht: (2025)
TransGAT: Transformer-Based Graph Neural Networks for Multi-Dimensional Automated Essay Scoring
von: Aljuaid, Hind, et al.
Veröffentlicht: (2025)
von: Aljuaid, Hind, et al.
Veröffentlicht: (2025)
Toward LLM-Supported Automated Assessment of Critical Thinking Subskills
von: Peczuh, Marisa C., et al.
Veröffentlicht: (2025)
von: Peczuh, Marisa C., et al.
Veröffentlicht: (2025)
ELMES: An Automated Framework for Evaluating Large Language Models in Educational Scenarios
von: Wei, Shou'ang, et al.
Veröffentlicht: (2025)
von: Wei, Shou'ang, et al.
Veröffentlicht: (2025)
Toward Automated Detection of Biased Social Signals from the Content of Clinical Conversations
von: Chen, Feng, et al.
Veröffentlicht: (2024)
von: Chen, Feng, et al.
Veröffentlicht: (2024)
CALM: Culturally Self-Aware Language Models
von: Shen, Lingzhi, et al.
Veröffentlicht: (2026)
von: Shen, Lingzhi, et al.
Veröffentlicht: (2026)
Reflecting in the Reflection: Integrating a Socratic Questioning Framework into Automated AI-Based Question Generation
von: Holub, Ondřej, et al.
Veröffentlicht: (2026)
von: Holub, Ondřej, et al.
Veröffentlicht: (2026)
TATA: Stance Detection via Topic-Agnostic and Topic-Aware Embeddings
von: Hanley, Hans W. A., et al.
Veröffentlicht: (2023)
von: Hanley, Hans W. A., et al.
Veröffentlicht: (2023)
Automating Governing Knowledge Commons and Contextual Integrity (GKC-CI) Privacy Policy Annotations with Large Language Models
von: Chanenson, Jake, et al.
Veröffentlicht: (2023)
von: Chanenson, Jake, et al.
Veröffentlicht: (2023)
The Perspectivist Paradigm Shift: Assumptions and Challenges of Capturing Human Labels
von: Fleisig, Eve, et al.
Veröffentlicht: (2024)
von: Fleisig, Eve, et al.
Veröffentlicht: (2024)
Hypothesis Testing for Quantifying LLM-Human Misalignment in Multiple Choice Settings
von: Hong, Harbin, et al.
Veröffentlicht: (2025)
von: Hong, Harbin, et al.
Veröffentlicht: (2025)
CycleResearcher: Improving Automated Research via Automated Review
von: Weng, Yixuan, et al.
Veröffentlicht: (2024)
von: Weng, Yixuan, et al.
Veröffentlicht: (2024)
KPoEM: A Human-Annotated Dataset for Emotion Classification and RAG-Based Poetry Generation in Korean Modern Poetry
von: Lim, Iro, et al.
Veröffentlicht: (2025)
von: Lim, Iro, et al.
Veröffentlicht: (2025)
Diagnosing Hate Speech Classification: Where Do Humans and Machines Disagree, and Why?
von: Yang, Xilin
Veröffentlicht: (2024)
von: Yang, Xilin
Veröffentlicht: (2024)
Does Continued Pretraining on a Learner Corpus Improve Automated Essay Scoring on English Proficiency Tests? Evidence from EFCAMDAT
von: Nguyen, Duy Anh
Veröffentlicht: (2026)
von: Nguyen, Duy Anh
Veröffentlicht: (2026)
Augmenting Human-Annotated Training Data with Large Language Model Generation and Distillation in Open-Response Assessment
von: Borchers, Conrad, et al.
Veröffentlicht: (2025)
von: Borchers, Conrad, et al.
Veröffentlicht: (2025)
Deep Learning Approaches for Detecting Adversarial Cyberbullying and Hate Speech in Social Networks
von: Azumah, Sylvia Worlali, et al.
Veröffentlicht: (2024)
von: Azumah, Sylvia Worlali, et al.
Veröffentlicht: (2024)
Vicarious Offense and Noise Audit of Offensive Speech Classifiers: Unifying Human and Machine Disagreement on What is Offensive
von: Weerasooriya, Tharindu Cyril, et al.
Veröffentlicht: (2023)
von: Weerasooriya, Tharindu Cyril, et al.
Veröffentlicht: (2023)
Accept or Deny? Evaluating LLM Fairness and Performance in Loan Approval across Table-to-Text Serialization Approaches
von: Azime, Israel Abebe, et al.
Veröffentlicht: (2025)
von: Azime, Israel Abebe, et al.
Veröffentlicht: (2025)
Thinking Outside the (Gray) Box: A Context-Based Score for Assessing Value and Originality in Neural Text Generation
von: Franceschelli, Giorgio, et al.
Veröffentlicht: (2025)
von: Franceschelli, Giorgio, et al.
Veröffentlicht: (2025)
Enhancing Automated Essay Scoring with Three Techniques: Two-Stage Fine-Tuning, Score Alignment, and Self-Training
von: Choi, Hongseok, et al.
Veröffentlicht: (2026)
von: Choi, Hongseok, et al.
Veröffentlicht: (2026)
Humanizing LLMs: A Survey of Psychological Measurements with Tools, Datasets, and Human-Agent Applications
von: Dong, Wenhan, et al.
Veröffentlicht: (2025)
von: Dong, Wenhan, et al.
Veröffentlicht: (2025)
Auto311: A Confidence-guided Automated System for Non-emergency Calls
von: Chen, Zirong, et al.
Veröffentlicht: (2023)
von: Chen, Zirong, et al.
Veröffentlicht: (2023)
Empowering Many, Biasing a Few: Generalist Credit Scoring through Large Language Models
von: Feng, Duanyu, et al.
Veröffentlicht: (2023)
von: Feng, Duanyu, et al.
Veröffentlicht: (2023)
MMM-fair: An Interactive Toolkit for Exploring and Operationalizing Multi-Fairness Trade-offs
von: Swati, Swati, et al.
Veröffentlicht: (2025)
von: Swati, Swati, et al.
Veröffentlicht: (2025)
Pair2Score: Pairwise-to-Absolute Transfer for LLM-Based Essay Scoring
von: Hallaç, İbrahim Rıza, et al.
Veröffentlicht: (2026)
von: Hallaç, İbrahim Rıza, et al.
Veröffentlicht: (2026)
Literature Meets Data: A Synergistic Approach to Hypothesis Generation
von: Liu, Haokun, et al.
Veröffentlicht: (2024)
von: Liu, Haokun, et al.
Veröffentlicht: (2024)
The Reasoning Trap -- Logical Reasoning as a Mechanistic Pathway to Situational Awareness
von: Sahoo, Subramanyam, et al.
Veröffentlicht: (2026)
von: Sahoo, Subramanyam, et al.
Veröffentlicht: (2026)
LLM Defenses Are Not Robust to Multi-Turn Human Jailbreaks Yet
von: Li, Nathaniel, et al.
Veröffentlicht: (2024)
von: Li, Nathaniel, et al.
Veröffentlicht: (2024)
A Practical Method for Generating String Counterfactuals
von: Avitan, Matan, et al.
Veröffentlicht: (2024)
von: Avitan, Matan, et al.
Veröffentlicht: (2024)
DetoxLLM: A Framework for Detoxification with Explanations
von: Khondaker, Md Tawkat Islam, et al.
Veröffentlicht: (2024)
von: Khondaker, Md Tawkat Islam, et al.
Veröffentlicht: (2024)
GRILE: A Benchmark for Grammar Reasoning and Explanation in Romanian LLMs
von: Dumitran, Adrian-Marius, et al.
Veröffentlicht: (2025)
von: Dumitran, Adrian-Marius, et al.
Veröffentlicht: (2025)
PRSM: A Measure to Evaluate CLIP's Robustness Against Paraphrases
von: Schlegel, Udo, et al.
Veröffentlicht: (2025)
von: Schlegel, Udo, et al.
Veröffentlicht: (2025)
Harnessing Large Language Models for Disaster Management: A Survey
von: Lei, Zhenyu, et al.
Veröffentlicht: (2025)
von: Lei, Zhenyu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
RDBE: Reasoning Distillation-Based Evaluation Enhances Automatic Essay Scoring
von: Mohammadkhani, Ali Ghiasvand
Veröffentlicht: (2024) -
Beyond the Score: Uncertainty-Calibrated LLMs for Automated Essay Assessment
von: Karim, Ahmed, et al.
Veröffentlicht: (2025) -
Exploration of Summarization by Generative Language Models for Automated Scoring of Long Essays
von: Hua, Haowei, et al.
Veröffentlicht: (2025) -
Is GPT-4 Alone Sufficient for Automated Essay Scoring?: A Comparative Judgment Approach Based on Rater Cognition
von: Kim, Seungju, et al.
Veröffentlicht: (2024) -
A Time-Aware Approach to Early Detection of Anorexia: UNSL at eRisk 2024
von: Thompson, Horacio, et al.
Veröffentlicht: (2024)