AERA Chat: An Interactive Platform for Automated Explainable Student Answer Assessment
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Jiazheng, Bobrov, Artem, Zhao, Runcong, Aloisi, Cesare, He, Yulan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
An Automated Explainable Educational Assessment System Built on LLMs
por: Li, Jiazheng, et al.
Publicado: (2024)
por: Li, Jiazheng, et al.
Publicado: (2024)
LearnLens: LLM-Enabled Personalised, Curriculum-Grounded Feedback with Educators in the Loop
por: Zhao, Runcong, et al.
Publicado: (2025)
por: Zhao, Runcong, et al.
Publicado: (2025)
Calibrating LLMs with Preference Optimization on Thought Trees for Generating Rationale in Science Question Scoring
por: Li, Jiazheng, et al.
Publicado: (2024)
por: Li, Jiazheng, et al.
Publicado: (2024)
Two Heads Are Better Than One: Dual-Model Verbal Reflection at Inference-Time
por: Li, Jiazheng, et al.
Publicado: (2025)
por: Li, Jiazheng, et al.
Publicado: (2025)
Large Language Models Fall Short: Understanding Complex Relationships in Detective Narratives
por: Zhao, Runcong, et al.
Publicado: (2024)
por: Zhao, Runcong, et al.
Publicado: (2024)
Are NLP Models Good at Tracing Thoughts: An Overview of Narrative Understanding
por: Zhu, Lixing, et al.
Publicado: (2023)
por: Zhu, Lixing, et al.
Publicado: (2023)
PLAYER*: Enhancing LLM-based Multi-Agent Communication and Interaction in Murder Mystery Games
por: Zhu, Qinglin, et al.
Publicado: (2024)
por: Zhu, Qinglin, et al.
Publicado: (2024)
ExDDI: Explaining Drug-Drug Interaction Predictions with Natural Language
por: Sun, Zhaoyue, et al.
Publicado: (2024)
por: Sun, Zhaoyue, et al.
Publicado: (2024)
OpenToM: A Comprehensive Benchmark for Evaluating Theory-of-Mind Reasoning Capabilities of Large Language Models
por: Xu, Hainiu, et al.
Publicado: (2024)
por: Xu, Hainiu, et al.
Publicado: (2024)
Soft Reasoning: Navigating Solution Spaces in Large Language Models through Controlled Embedding Exploration
por: Zhu, Qinglin, et al.
Publicado: (2025)
por: Zhu, Qinglin, et al.
Publicado: (2025)
SymbolicThought: Integrating Language Models and Symbolic Reasoning for Consistent and Interpretable Human Relationship Understanding
por: Zhao, Runcong, et al.
Publicado: (2025)
por: Zhao, Runcong, et al.
Publicado: (2025)
Beyond RAG for Agent Memory: Retrieval by Decoupling and Aggregation
por: Hu, Zhanghao, et al.
Publicado: (2026)
por: Hu, Zhanghao, et al.
Publicado: (2026)
Eliminating Biased Length Reliance of Direct Preference Optimization via Down-Sampled KL Divergence
por: Lu, Junru, et al.
Publicado: (2024)
por: Lu, Junru, et al.
Publicado: (2024)
Detecting Contextual Hallucinations in LLMs with Frequency-Aware Attention
por: Qi, Siya, et al.
Publicado: (2026)
por: Qi, Siya, et al.
Publicado: (2026)
The Mystery of In-Context Learning: A Comprehensive Survey on Interpretation and Analysis
por: Zhou, Yuxiang, et al.
Publicado: (2023)
por: Zhou, Yuxiang, et al.
Publicado: (2023)
Sparse Activation Editing for Reliable Instruction Following in Narratives
por: Zhao, Runcong, et al.
Publicado: (2025)
por: Zhao, Runcong, et al.
Publicado: (2025)
Leveraging ChatGPT in Pharmacovigilance Event Extraction: An Empirical Study
por: Sun, Zhaoyue, et al.
Publicado: (2024)
por: Sun, Zhaoyue, et al.
Publicado: (2024)
DrugWatch: A Comprehensive Multi-Source Data Visualisation Platform for Drug Safety Information
por: Bobrov, Artem, et al.
Publicado: (2024)
por: Bobrov, Artem, et al.
Publicado: (2024)
Focusing on Students, not Machines: Grounded Question Generation and Automated Answer Grading
por: Meyer, Gérôme, et al.
Publicado: (2025)
por: Meyer, Gérôme, et al.
Publicado: (2025)
Explainable Depression Detection in Clinical Interviews with Personalized Retrieval-Augmented Generation
por: Zhang, Linhai, et al.
Publicado: (2025)
por: Zhang, Linhai, et al.
Publicado: (2025)
GraphMind: Interactive Novelty Assessment System for Accelerating Scientific Discovery
por: da Silva, Italo Luis, et al.
Publicado: (2025)
por: da Silva, Italo Luis, et al.
Publicado: (2025)
Latent Refinement Decoding: Enhancing Diffusion-Based Language Models by Refining Belief States
por: Zhu, Qinglin, et al.
Publicado: (2025)
por: Zhu, Qinglin, et al.
Publicado: (2025)
VegaChat: A Robust Framework for LLM-Based Chart Generation and Assessment
por: Hostnik, Marko, et al.
Publicado: (2026)
por: Hostnik, Marko, et al.
Publicado: (2026)
SSA: Sparse Sparse Attention by Aligning Full and Sparse Attention Outputs in Feature Space
por: Shen, Zhenyi, et al.
Publicado: (2025)
por: Shen, Zhenyi, et al.
Publicado: (2025)
EnigmaToM: Improve LLMs' Theory-of-Mind Reasoning Capabilities with Neural Knowledge Base of Entity States
por: Xu, Hainiu, et al.
Publicado: (2025)
por: Xu, Hainiu, et al.
Publicado: (2025)
Explainable Recommender with Geometric Information Bottleneck
por: Yan, Hanqi, et al.
Publicado: (2023)
por: Yan, Hanqi, et al.
Publicado: (2023)
RoleMRC: A Fine-Grained Composite Benchmark for Role-Playing and Instruction-Following
por: Lu, Junru, et al.
Publicado: (2025)
por: Lu, Junru, et al.
Publicado: (2025)
Linear Ensembles Wash Away Watermarks: On the Fragility of Distributional Perturbations in LLMs
por: Wu, Zhihao, et al.
Publicado: (2026)
por: Wu, Zhihao, et al.
Publicado: (2026)
WildChat: 1M ChatGPT Interaction Logs in the Wild
por: Zhao, Wenting, et al.
Publicado: (2024)
por: Zhao, Wenting, et al.
Publicado: (2024)
AgentMental: An Interactive Multi-Agent Framework for Explainable and Adaptive Mental Health Assessment
por: Hu, Jinpeng, et al.
Publicado: (2025)
por: Hu, Jinpeng, et al.
Publicado: (2025)
Evaluating ChatGPT on Medical Information Extraction Tasks: Performance, Explainability and Beyond
por: Li, Liz, et al.
Publicado: (2026)
por: Li, Liz, et al.
Publicado: (2026)
RECIPE4U: Student-ChatGPT Interaction Dataset in EFL Writing Education
por: Han, Jieun, et al.
Publicado: (2024)
por: Han, Jieun, et al.
Publicado: (2024)
Automated Long Answer Grading with RiceChem Dataset
por: Sonkar, Shashank, et al.
Publicado: (2024)
por: Sonkar, Shashank, et al.
Publicado: (2024)
Employing Label Models on ChatGPT Answers Improves Legal Text Entailment Performance
por: Nguyen, Chau, et al.
Publicado: (2024)
por: Nguyen, Chau, et al.
Publicado: (2024)
Student Answer Forecasting: Transformer-Driven Answer Choice Prediction for Language Learning
por: Gado, Elena Grazia, et al.
Publicado: (2024)
por: Gado, Elena Grazia, et al.
Publicado: (2024)
Domaino1s: Guiding LLM Reasoning for Explainable Answers in High-Stakes Domains
por: Chu, Xu, et al.
Publicado: (2025)
por: Chu, Xu, et al.
Publicado: (2025)
Pull Requests as a Training Signal for Repo-Level Code Editing
por: Zhu, Qinglin, et al.
Publicado: (2026)
por: Zhu, Qinglin, et al.
Publicado: (2026)
BioACE: An Automated Framework for Biomedical Answer and Citation Evaluations
por: Gupta, Deepak, et al.
Publicado: (2026)
por: Gupta, Deepak, et al.
Publicado: (2026)
Automated Assessment of Students' Code Comprehension using LLMs
por: Oli, Priti, et al.
Publicado: (2023)
por: Oli, Priti, et al.
Publicado: (2023)
Automated Answer Validation using Text Similarity
por: Ganesan, Balaji, et al.
Publicado: (2024)
por: Ganesan, Balaji, et al.
Publicado: (2024)
Ejemplares similares
-
An Automated Explainable Educational Assessment System Built on LLMs
por: Li, Jiazheng, et al.
Publicado: (2024) -
LearnLens: LLM-Enabled Personalised, Curriculum-Grounded Feedback with Educators in the Loop
por: Zhao, Runcong, et al.
Publicado: (2025) -
Calibrating LLMs with Preference Optimization on Thought Trees for Generating Rationale in Science Question Scoring
por: Li, Jiazheng, et al.
Publicado: (2024) -
Two Heads Are Better Than One: Dual-Model Verbal Reflection at Inference-Time
por: Li, Jiazheng, et al.
Publicado: (2025) -
Large Language Models Fall Short: Understanding Complex Relationships in Detective Narratives
por: Zhao, Runcong, et al.
Publicado: (2024)