An Automated Explainable Educational Assessment System Built on LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Jiazheng, Bobrov, Artem, West, David, Aloisi, Cesare, He, Yulan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AERA Chat: An Interactive Platform for Automated Explainable Student Answer Assessment
by: Li, Jiazheng, et al.
Published: (2024)
by: Li, Jiazheng, et al.
Published: (2024)
LearnLens: LLM-Enabled Personalised, Curriculum-Grounded Feedback with Educators in the Loop
by: Zhao, Runcong, et al.
Published: (2025)
by: Zhao, Runcong, et al.
Published: (2025)
Calibrating LLMs with Preference Optimization on Thought Trees for Generating Rationale in Science Question Scoring
by: Li, Jiazheng, et al.
Published: (2024)
by: Li, Jiazheng, et al.
Published: (2024)
Two Heads Are Better Than One: Dual-Model Verbal Reflection at Inference-Time
by: Li, Jiazheng, et al.
Published: (2025)
by: Li, Jiazheng, et al.
Published: (2025)
ExDDI: Explaining Drug-Drug Interaction Predictions with Natural Language
by: Sun, Zhaoyue, et al.
Published: (2024)
by: Sun, Zhaoyue, et al.
Published: (2024)
EnigmaToM: Improve LLMs' Theory-of-Mind Reasoning Capabilities with Neural Knowledge Base of Entity States
by: Xu, Hainiu, et al.
Published: (2025)
by: Xu, Hainiu, et al.
Published: (2025)
The Mystery of In-Context Learning: A Comprehensive Survey on Interpretation and Analysis
by: Zhou, Yuxiang, et al.
Published: (2023)
by: Zhou, Yuxiang, et al.
Published: (2023)
Explainable Depression Detection in Clinical Interviews with Personalized Retrieval-Augmented Generation
by: Zhang, Linhai, et al.
Published: (2025)
by: Zhang, Linhai, et al.
Published: (2025)
Eliminating Biased Length Reliance of Direct Preference Optimization via Down-Sampled KL Divergence
by: Lu, Junru, et al.
Published: (2024)
by: Lu, Junru, et al.
Published: (2024)
Evaluating LLMs' Assessment of Mixed-Context Hallucination Through the Lens of Summarization
by: Qi, Siya, et al.
Published: (2025)
by: Qi, Siya, et al.
Published: (2025)
SSA: Sparse Sparse Attention by Aligning Full and Sparse Attention Outputs in Feature Space
by: Shen, Zhenyi, et al.
Published: (2025)
by: Shen, Zhenyi, et al.
Published: (2025)
Explainable Recommender with Geometric Information Bottleneck
by: Yan, Hanqi, et al.
Published: (2023)
by: Yan, Hanqi, et al.
Published: (2023)
RoleMRC: A Fine-Grained Composite Benchmark for Role-Playing and Instruction-Following
by: Lu, Junru, et al.
Published: (2025)
by: Lu, Junru, et al.
Published: (2025)
GraphMind: Interactive Novelty Assessment System for Accelerating Scientific Discovery
by: da Silva, Italo Luis, et al.
Published: (2025)
by: da Silva, Italo Luis, et al.
Published: (2025)
Large Language Models Fall Short: Understanding Complex Relationships in Detective Narratives
by: Zhao, Runcong, et al.
Published: (2024)
by: Zhao, Runcong, et al.
Published: (2024)
Explainability and Certification of AI-Generated Educational Assessments
by: Yaacoub, Antoun, et al.
Published: (2026)
by: Yaacoub, Antoun, et al.
Published: (2026)
Using LLMs for Automated Privacy Policy Analysis: Prompt Engineering, Fine-Tuning and Explainability
by: Chen, Yuxin, et al.
Published: (2025)
by: Chen, Yuxin, et al.
Published: (2025)
Multi-dimensional Assessment and Explainable Feedback for Counselor Responses to Client Resistance in Text-based Counseling with LLMs
by: Li, Anqi, et al.
Published: (2026)
by: Li, Anqi, et al.
Published: (2026)
Automated Assessment of Students' Code Comprehension using LLMs
by: Oli, Priti, et al.
Published: (2023)
by: Oli, Priti, et al.
Published: (2023)
Assessing the Reasoning Capabilities of LLMs in the context of Evidence-based Claim Verification
by: Dougrez-Lewis, John, et al.
Published: (2024)
by: Dougrez-Lewis, John, et al.
Published: (2024)
Do LLMs Give Psychometrically Plausible Responses in Educational Assessments?
by: Säuberli, Andreas, et al.
Published: (2025)
by: Säuberli, Andreas, et al.
Published: (2025)
LLMs for Explainable AI: A Comprehensive Survey
by: Bilal, Ahsan, et al.
Published: (2025)
by: Bilal, Ahsan, et al.
Published: (2025)
Machine Translation in the Wild: User Reaction to Xiaohongshu's Built-In Translation Feature
by: He, Sui
Published: (2026)
by: He, Sui
Published: (2026)
Extracting Event Temporal Relations via Hyperbolic Geometry
by: Tan, Xingwei, et al.
Published: (2021)
by: Tan, Xingwei, et al.
Published: (2021)
Beyond the Score: Uncertainty-Calibrated LLMs for Automated Essay Assessment
by: Karim, Ahmed, et al.
Published: (2025)
by: Karim, Ahmed, et al.
Published: (2025)
Weak Reward Model Transforms Generative Models into Robust Causal Event Extraction Systems
by: da Silva, Italo Luis, et al.
Published: (2024)
by: da Silva, Italo Luis, et al.
Published: (2024)
DrugWatch: A Comprehensive Multi-Source Data Visualisation Platform for Drug Safety Information
by: Bobrov, Artem, et al.
Published: (2024)
by: Bobrov, Artem, et al.
Published: (2024)
Training and Evaluation of Guideline-Based Medical Reasoning in LLMs
by: Staniek, Michael, et al.
Published: (2025)
by: Staniek, Michael, et al.
Published: (2025)
From Black Box to Transparency: Enhancing Automated Interpreting Assessment with Explainable AI in College Classrooms
by: Jiang, Zhaokun, et al.
Published: (2025)
by: Jiang, Zhaokun, et al.
Published: (2025)
Semantic Refinement with LLMs for Graph Representations
by: Thapaliya, Safal, et al.
Published: (2025)
by: Thapaliya, Safal, et al.
Published: (2025)
Detecting Contextual Hallucinations in LLMs with Frequency-Aware Attention
by: Qi, Siya, et al.
Published: (2026)
by: Qi, Siya, et al.
Published: (2026)
Interpretability from the Ground Up: Stakeholder-Centric Design of Automated Scoring in Educational Assessments
by: Kim, Yunsung, et al.
Published: (2025)
by: Kim, Yunsung, et al.
Published: (2025)
Automated Bias Assessment in AI-Generated Educational Content Using CEAT Framework
by: Peng, Jingyang, et al.
Published: (2025)
by: Peng, Jingyang, et al.
Published: (2025)
Towards Self-Improving Error Diagnosis in Multi-Agent Systems
by: Li, Jiazheng, et al.
Published: (2026)
by: Li, Jiazheng, et al.
Published: (2026)
Towards Dog Bark Decoding: Leveraging Human Speech Processing for Automated Bark Classification
by: Abzaliev, Artem, et al.
Published: (2024)
by: Abzaliev, Artem, et al.
Published: (2024)
Triadic Fusion of Cognitive, Functional, and Causal Dimensions for Explainable LLMs: The TAXAL Framework
by: Herrera-Poyatos, David, et al.
Published: (2025)
by: Herrera-Poyatos, David, et al.
Published: (2025)
CRAVE: A Conflicting Reasoning Approach for Explainable Claim Verification Using LLMs
by: Zheng, Yingming, et al.
Published: (2025)
by: Zheng, Yingming, et al.
Published: (2025)
Rehearse With User: Personalized Opinion Summarization via Role-Playing based on Large Language Models
by: Zhang, Yanyue, et al.
Published: (2025)
by: Zhang, Yanyue, et al.
Published: (2025)
CWTM: Leveraging Contextualized Word Embeddings from BERT for Neural Topic Modeling
by: Fang, Zheng, et al.
Published: (2023)
by: Fang, Zheng, et al.
Published: (2023)
HyLaT: Efficient Multi-Agent Communication via Hybrid Latent-Text Protocol
by: Mou, Xinyi, et al.
Published: (2026)
by: Mou, Xinyi, et al.
Published: (2026)
Similar Items
-
AERA Chat: An Interactive Platform for Automated Explainable Student Answer Assessment
by: Li, Jiazheng, et al.
Published: (2024) -
LearnLens: LLM-Enabled Personalised, Curriculum-Grounded Feedback with Educators in the Loop
by: Zhao, Runcong, et al.
Published: (2025) -
Calibrating LLMs with Preference Optimization on Thought Trees for Generating Rationale in Science Question Scoring
by: Li, Jiazheng, et al.
Published: (2024) -
Two Heads Are Better Than One: Dual-Model Verbal Reflection at Inference-Time
by: Li, Jiazheng, et al.
Published: (2025) -
ExDDI: Explaining Drug-Drug Interaction Predictions with Natural Language
by: Sun, Zhaoyue, et al.
Published: (2024)