EssayCBM: Rubric-Aligned Concept Bottleneck Models for Transparent Essay Grading
Fuente:
arXiv
Saved in:
| Main Authors: | Chaudhary, Kumar Satvik, Zhao, Chengshuai, Zhang, Fan, Agrawal, Garima, Deng, Yuli, Liu, Huan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
REC-CBM: Rubric-Aware Error-Correction Concept Bottleneck Models for Trustworthy Open-Ended Grading
by: Zhao, Chengshuai, et al.
Published: (2026)
by: Zhao, Chengshuai, et al.
Published: (2026)
Hey AI Can You Grade My Essay?: Automatic Essay Grading
by: Maliha, Maisha, et al.
Published: (2024)
by: Maliha, Maisha, et al.
Published: (2024)
Domain Knowledge-Enhanced LLMs for Fraud and Concept Drift Detection
by: Şenol, Ali, et al.
Published: (2025)
by: Şenol, Ali, et al.
Published: (2025)
Automated Refinement of Essay Scoring Rubrics for Language Models via Reflect-and-Revise
by: Harada, Keno, et al.
Published: (2025)
by: Harada, Keno, et al.
Published: (2025)
TRATES: Trait-Specific Rubric-Assisted Cross-Prompt Essay Scoring
by: Eltanbouly, Sohaila, et al.
Published: (2025)
by: Eltanbouly, Sohaila, et al.
Published: (2025)
DREsS: Dataset for Rubric-based Essay Scoring on EFL Writing
by: Yoo, Haneul, et al.
Published: (2024)
by: Yoo, Haneul, et al.
Published: (2024)
Joint Detection of Fraud and Concept Drift inOnline Conversations with LLM-Assisted Judgment
by: Senol, Ali, et al.
Published: (2025)
by: Senol, Ali, et al.
Published: (2025)
LLMs Do Not Grade Essays Like Humans
by: Mathew, Jerin George, et al.
Published: (2026)
by: Mathew, Jerin George, et al.
Published: (2026)
LLM Essay Scoring Under Holistic and Analytic Rubrics: Prompt Effects and Bias
by: Kucia, Filip J., et al.
Published: (2026)
by: Kucia, Filip J., et al.
Published: (2026)
Do We Need a Detailed Rubric for Automated Essay Scoring using Large Language Models?
by: Yoshida, Lui
Published: (2025)
by: Yoshida, Lui
Published: (2025)
How well can LLMs Grade Essays in Arabic?
by: Ghazawi, Rayed, et al.
Published: (2025)
by: Ghazawi, Rayed, et al.
Published: (2025)
EssayBench: Evaluating Large Language Models in Multi-Genre Chinese Essay Writing
by: Gao, Fan, et al.
Published: (2025)
by: Gao, Fan, et al.
Published: (2025)
Ontology-Aware RAG for Improved Question-Answering in Cybersecurity Education
by: Zhao, Chengshuai, et al.
Published: (2024)
by: Zhao, Chengshuai, et al.
Published: (2024)
FeedEval: Pedagogically Aligned Evaluation of LLM-Generated Essay Feedback
by: Chu, Seongyeub, et al.
Published: (2026)
by: Chu, Seongyeub, et al.
Published: (2026)
CLIF: Concept-Level Influence Functions for Transparent Bottleneck Models
by: Sun, Yike, et al.
Published: (2026)
by: Sun, Yike, et al.
Published: (2026)
Measuring Reasoning Quality in LLMs: A Multi-Dimensional Behavioral Framework
by: Şenol, Ali, et al.
Published: (2026)
by: Şenol, Ali, et al.
Published: (2026)
Specialists or Generalists? Multi-Agent and Single-Agent LLMs for Essay Grading
by: Idowu, Jamiu Adekunle, et al.
Published: (2026)
by: Idowu, Jamiu Adekunle, et al.
Published: (2026)
Are Today's LLMs Ready to Explain Well-Being Concepts?
by: Jiang, Bohan, et al.
Published: (2025)
by: Jiang, Bohan, et al.
Published: (2025)
Can Large Language Models Differentiate Harmful from Argumentative Essays? Steps Toward Ethical Essay Scoring
by: Kim, Hongjin, et al.
Published: (2026)
by: Kim, Hongjin, et al.
Published: (2026)
EssayJudge: A Multi-Granular Benchmark for Assessing Automated Essay Scoring Capabilities of Multimodal Large Language Models
by: Su, Jiamin, et al.
Published: (2025)
by: Su, Jiamin, et al.
Published: (2025)
Evaluating Austrian A-Level German Essays with Large Language Models for Automated Essay Scoring
by: Kubesch, Jonas, et al.
Published: (2026)
by: Kubesch, Jonas, et al.
Published: (2026)
Machine-Assisted Grading of Nationwide School-Leaving Essay Exams with LLMs and Statistical NLP
by: Karjus, Andres, et al.
Published: (2026)
by: Karjus, Andres, et al.
Published: (2026)
CyberBOT: Towards Reliable Cybersecurity Education via Ontology-Grounded Retrieval Augmented Generation
by: Zhao, Chengshuai, et al.
Published: (2025)
by: Zhao, Chengshuai, et al.
Published: (2025)
Composable Cross-prompt Essay Scoring by Merging Models
by: Lee, Sanwoo, et al.
Published: (2025)
by: Lee, Sanwoo, et al.
Published: (2025)
Long Context Automated Essay Scoring with Language Models
by: Ormerod, Christopher, et al.
Published: (2025)
by: Ormerod, Christopher, et al.
Published: (2025)
Rubric-Conditioned LLM Grading: Alignment, Uncertainty, and Robustness
by: Deng, Haotian, et al.
Published: (2025)
by: Deng, Haotian, et al.
Published: (2025)
Are Large Language Models Good Essay Graders?
by: Kundu, Anindita, et al.
Published: (2024)
by: Kundu, Anindita, et al.
Published: (2024)
Can Knowledge Graphs Reduce Hallucinations in LLMs? : A Survey
by: Agrawal, Garima, et al.
Published: (2023)
by: Agrawal, Garima, et al.
Published: (2023)
Catching Chameleons: Detecting Evolving Disinformation Generated using Large Language Models
by: Jiang, Bohan, et al.
Published: (2024)
by: Jiang, Bohan, et al.
Published: (2024)
Adversarial Topic-aware Prompt-tuning for Cross-topic Automated Essay Scoring
by: Zhang, Chunyun, et al.
Published: (2025)
by: Zhang, Chunyun, et al.
Published: (2025)
Rationale Behind Essay Scores: Enhancing S-LLM's Multi-Trait Essay Scoring with Rationale Generated by LLMs
by: Chu, SeongYeub, et al.
Published: (2024)
by: Chu, SeongYeub, et al.
Published: (2024)
Automatic Essay Scoring in a Brazilian Scenario
by: Matsuoka, Felipe Akio
Published: (2023)
by: Matsuoka, Felipe Akio
Published: (2023)
Implicit Grading Bias in Large Language Models: How Writing Style Affects Automated Assessment Across Math, Programming, and Essay Tasks
by: Jadhav, Rudra, et al.
Published: (2026)
by: Jadhav, Rudra, et al.
Published: (2026)
Decision-Level Ordinal Modeling for Multimodal Essay Scoring with Large Language Models
by: Zhang, Han, et al.
Published: (2026)
by: Zhang, Han, et al.
Published: (2026)
Unleashing Large Language Models' Proficiency in Zero-shot Essay Scoring
by: Lee, Sanwoo, et al.
Published: (2024)
by: Lee, Sanwoo, et al.
Published: (2024)
Context is Enough: Empirical Validation of $\textit{Sequentiality}$ on Essays
by: Sunny, Amal, et al.
Published: (2025)
by: Sunny, Amal, et al.
Published: (2025)
Enhancing Essay Scoring with Adversarial Weights Perturbation and Metric-specific AttentionPooling
by: Huang, Jiaxin, et al.
Published: (2024)
by: Huang, Jiaxin, et al.
Published: (2024)
V2C-CBM: Building Concept Bottlenecks with Vision-to-Concept Tokenizer
by: He, Hangzhou, et al.
Published: (2025)
by: He, Hangzhou, et al.
Published: (2025)
AdaCBM: An Adaptive Concept Bottleneck Model for Explainable and Accurate Diagnosis
by: Chowdhury, Townim F., et al.
Published: (2024)
by: Chowdhury, Townim F., et al.
Published: (2024)
VLG-CBM: Training Concept Bottleneck Models with Vision-Language Guidance
by: Srivastava, Divyansh, et al.
Published: (2024)
by: Srivastava, Divyansh, et al.
Published: (2024)
Similar Items
-
REC-CBM: Rubric-Aware Error-Correction Concept Bottleneck Models for Trustworthy Open-Ended Grading
by: Zhao, Chengshuai, et al.
Published: (2026) -
Hey AI Can You Grade My Essay?: Automatic Essay Grading
by: Maliha, Maisha, et al.
Published: (2024) -
Domain Knowledge-Enhanced LLMs for Fraud and Concept Drift Detection
by: Şenol, Ali, et al.
Published: (2025) -
Automated Refinement of Essay Scoring Rubrics for Language Models via Reflect-and-Revise
by: Harada, Keno, et al.
Published: (2025) -
TRATES: Trait-Specific Rubric-Assisted Cross-Prompt Essay Scoring
by: Eltanbouly, Sohaila, et al.
Published: (2025)