LCES: Zero-shot Automated Essay Scoring via Pairwise Comparisons Using Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shibata, Takumi, Miyamura, Yuichi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Rank-Then-Score: Enhancing Large Language Models for Automated Essay Scoring
von: Cai, Yida, et al.
Veröffentlicht: (2025)
von: Cai, Yida, et al.
Veröffentlicht: (2025)
Evaluating Austrian A-Level German Essays with Large Language Models for Automated Essay Scoring
von: Kubesch, Jonas, et al.
Veröffentlicht: (2026)
von: Kubesch, Jonas, et al.
Veröffentlicht: (2026)
EssayJudge: A Multi-Granular Benchmark for Assessing Automated Essay Scoring Capabilities of Multimodal Large Language Models
von: Su, Jiamin, et al.
Veröffentlicht: (2025)
von: Su, Jiamin, et al.
Veröffentlicht: (2025)
Pair2Score: Pairwise-to-Absolute Transfer for LLM-Based Essay Scoring
von: Hallaç, İbrahim Rıza, et al.
Veröffentlicht: (2026)
von: Hallaç, İbrahim Rıza, et al.
Veröffentlicht: (2026)
Can Large Language Models Automatically Score Proficiency of Written Essays?
von: Mansour, Watheq, et al.
Veröffentlicht: (2024)
von: Mansour, Watheq, et al.
Veröffentlicht: (2024)
Decision-Level Ordinal Modeling for Multimodal Essay Scoring with Large Language Models
von: Zhang, Han, et al.
Veröffentlicht: (2026)
von: Zhang, Han, et al.
Veröffentlicht: (2026)
Estimating the Error of Large Language Models at Pairwise Text Comparison
von: Li, Tianyi
Veröffentlicht: (2025)
von: Li, Tianyi
Veröffentlicht: (2025)
Automated Essay Scoring Incorporating Annotations from Automated Feedback Systems
von: Ormerod, Christopher
Veröffentlicht: (2025)
von: Ormerod, Christopher
Veröffentlicht: (2025)
ZeroDL: Zero-shot Distribution Learning for Text Clustering via Large Language Models
von: Jo, Hwiyeol, et al.
Veröffentlicht: (2024)
von: Jo, Hwiyeol, et al.
Veröffentlicht: (2024)
AI-generated Essays: Characteristics and Implications on Automated Scoring and Academic Integrity
von: Zhong, Yang, et al.
Veröffentlicht: (2024)
von: Zhong, Yang, et al.
Veröffentlicht: (2024)
Knockout LLM Assessment: Using Large Language Models for Evaluations through Iterative Pairwise Comparisons
von: Sandan, Isik Baran, et al.
Veröffentlicht: (2025)
von: Sandan, Isik Baran, et al.
Veröffentlicht: (2025)
Towards Prompt Generalization: Grammar-aware Cross-Prompt Automated Essay Scoring
von: Do, Heejin, et al.
Veröffentlicht: (2025)
von: Do, Heejin, et al.
Veröffentlicht: (2025)
Unleashing Large Language Models' Proficiency in Zero-shot Essay Scoring
von: Lee, Sanwoo, et al.
Veröffentlicht: (2024)
von: Lee, Sanwoo, et al.
Veröffentlicht: (2024)
MedAgents: Large Language Models as Collaborators for Zero-shot Medical Reasoning
von: Tang, Xiangru, et al.
Veröffentlicht: (2023)
von: Tang, Xiangru, et al.
Veröffentlicht: (2023)
Direct-Scoring NLG Evaluators Can Use Pairwise Comparisons Too
von: Lawrence, Logan, et al.
Veröffentlicht: (2025)
von: Lawrence, Logan, et al.
Veröffentlicht: (2025)
EssayBench: Evaluating Large Language Models in Multi-Genre Chinese Essay Writing
von: Gao, Fan, et al.
Veröffentlicht: (2025)
von: Gao, Fan, et al.
Veröffentlicht: (2025)
Automatic Essay Scoring and Feedback Generation in Basque Language Learning
von: Azurmendi, Ekhi, et al.
Veröffentlicht: (2025)
von: Azurmendi, Ekhi, et al.
Veröffentlicht: (2025)
Evaluation of the Automated Labeling Method for Taxonomic Nomenclature Through Prompt-Optimized Large Language Model
von: Inoshita, Keito, et al.
Veröffentlicht: (2025)
von: Inoshita, Keito, et al.
Veröffentlicht: (2025)
Autoregressive Score Generation for Multi-trait Essay Scoring
von: Do, Heejin, et al.
Veröffentlicht: (2024)
von: Do, Heejin, et al.
Veröffentlicht: (2024)
Rationale Behind Essay Scores: Enhancing S-LLM's Multi-Trait Essay Scoring with Rationale Generated by LLMs
von: Chu, SeongYeub, et al.
Veröffentlicht: (2024)
von: Chu, SeongYeub, et al.
Veröffentlicht: (2024)
Attention-guided Self-reflection for Zero-shot Hallucination Detection in Large Language Models
von: Liu, Qiang, et al.
Veröffentlicht: (2025)
von: Liu, Qiang, et al.
Veröffentlicht: (2025)
Large Language Models as Zero-shot Dialogue State Tracker through Function Calling
von: Li, Zekun, et al.
Veröffentlicht: (2024)
von: Li, Zekun, et al.
Veröffentlicht: (2024)
ZeFaV: Boosting Large Language Models for Zero-shot Fact Verification
von: Luu, Son T., et al.
Veröffentlicht: (2024)
von: Luu, Son T., et al.
Veröffentlicht: (2024)
Are Large Language Models Good Essay Graders?
von: Kundu, Anindita, et al.
Veröffentlicht: (2024)
von: Kundu, Anindita, et al.
Veröffentlicht: (2024)
Autoregressive Multi-trait Essay Scoring via Reinforcement Learning with Scoring-aware Multiple Rewards
von: Do, Heejin, et al.
Veröffentlicht: (2024)
von: Do, Heejin, et al.
Veröffentlicht: (2024)
Evaluating Large Language Models on the Frame and Symbol Grounding Problems: A Zero-shot Benchmark
von: Oka, Shoko
Veröffentlicht: (2025)
von: Oka, Shoko
Veröffentlicht: (2025)
AutoSCORE: Enhancing Automated Scoring with Multi-Agent Large Language Models via Structured Component Recognition
von: Wang, Yun, et al.
Veröffentlicht: (2025)
von: Wang, Yun, et al.
Veröffentlicht: (2025)
Zero-shot and Few-shot Learning with Instruction-following LLMs for Claim Matching in Automated Fact-checking
von: Pisarevskaya, Dina, et al.
Veröffentlicht: (2025)
von: Pisarevskaya, Dina, et al.
Veröffentlicht: (2025)
Teach-to-Reason with Scoring: Self-Explainable Rationale-Driven Multi-Trait Essay Scoring
von: Do, Heejin, et al.
Veröffentlicht: (2025)
von: Do, Heejin, et al.
Veröffentlicht: (2025)
A Zero-shot and Few-shot Study of Instruction-Finetuned Large Language Models Applied to Clinical and Biomedical Tasks
von: Labrak, Yanis, et al.
Veröffentlicht: (2023)
von: Labrak, Yanis, et al.
Veröffentlicht: (2023)
Automating Adjudication of Cardiovascular Events Using Large Language Models
von: Sivarajkumar, Sonish, et al.
Veröffentlicht: (2025)
von: Sivarajkumar, Sonish, et al.
Veröffentlicht: (2025)
Activations as Features: Probing LLMs for Generalizable Essay Scoring Representations
von: Chi, Jinwei, et al.
Veröffentlicht: (2025)
von: Chi, Jinwei, et al.
Veröffentlicht: (2025)
Improve LLM-based Automatic Essay Scoring with Linguistic Features
von: Hou, Zhaoyi Joey, et al.
Veröffentlicht: (2025)
von: Hou, Zhaoyi Joey, et al.
Veröffentlicht: (2025)
DREsS: Dataset for Rubric-based Essay Scoring on EFL Writing
von: Yoo, Haneul, et al.
Veröffentlicht: (2024)
von: Yoo, Haneul, et al.
Veröffentlicht: (2024)
DeCAP: Context-Adaptive Prompt Generation for Debiasing Zero-shot Question Answering in Large Language Models
von: Bae, Suyoung, et al.
Veröffentlicht: (2025)
von: Bae, Suyoung, et al.
Veröffentlicht: (2025)
Zero-shot Benchmarking: A Framework for Flexible and Scalable Automatic Evaluation of Language Models
von: Pombal, José, et al.
Veröffentlicht: (2025)
von: Pombal, José, et al.
Veröffentlicht: (2025)
SpatialPrompting: Keyframe-driven Zero-Shot Spatial Reasoning with Off-the-Shelf Multimodal Large Language Models
von: Taguchi, Shun, et al.
Veröffentlicht: (2025)
von: Taguchi, Shun, et al.
Veröffentlicht: (2025)
Automatic Essay Multi-dimensional Scoring with Fine-tuning and Multiple Regression
von: Sun, Kun, et al.
Veröffentlicht: (2024)
von: Sun, Kun, et al.
Veröffentlicht: (2024)
Med-V1: Small Language Models for Zero-shot and Scalable Biomedical Evidence Attribution
von: Jin, Qiao, et al.
Veröffentlicht: (2026)
von: Jin, Qiao, et al.
Veröffentlicht: (2026)
Online Rubrics Elicitation from Pairwise Comparisons
von: Rezaei, MohammadHossein, et al.
Veröffentlicht: (2025)
von: Rezaei, MohammadHossein, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Rank-Then-Score: Enhancing Large Language Models for Automated Essay Scoring
von: Cai, Yida, et al.
Veröffentlicht: (2025) -
Evaluating Austrian A-Level German Essays with Large Language Models for Automated Essay Scoring
von: Kubesch, Jonas, et al.
Veröffentlicht: (2026) -
EssayJudge: A Multi-Granular Benchmark for Assessing Automated Essay Scoring Capabilities of Multimodal Large Language Models
von: Su, Jiamin, et al.
Veröffentlicht: (2025) -
Pair2Score: Pairwise-to-Absolute Transfer for LLM-Based Essay Scoring
von: Hallaç, İbrahim Rıza, et al.
Veröffentlicht: (2026) -
Can Large Language Models Automatically Score Proficiency of Written Essays?
von: Mansour, Watheq, et al.
Veröffentlicht: (2024)