Beyond Scores: A Modular RAG-Based System for Automatic Short Answer Scoring with Feedback
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fateen, Menna, Wang, Bo, Mine, Tsunenori |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Developing a Tutoring Dialog Dataset to Optimize LLMs for Educational Use
von: Fateen, Menna, et al.
Veröffentlicht: (2024)
von: Fateen, Menna, et al.
Veröffentlicht: (2024)
One Stone, Four Birds: A Comprehensive Solution for QA System Using Supervised Contrastive Learning
von: Wang, Bo, et al.
Veröffentlicht: (2024)
von: Wang, Bo, et al.
Veröffentlicht: (2024)
Automatic Essay Scoring and Feedback Generation in Basque Language Learning
von: Azurmendi, Ekhi, et al.
Veröffentlicht: (2025)
von: Azurmendi, Ekhi, et al.
Veröffentlicht: (2025)
Generative Language Models with Retrieval Augmented Generation for Automated Short Answer Scoring
von: Wang, Zifan, et al.
Veröffentlicht: (2024)
von: Wang, Zifan, et al.
Veröffentlicht: (2024)
SAS-Bench: A Fine-Grained Benchmark for Evaluating Short Answer Scoring with Large Language Models
von: Lai, Peichao, et al.
Veröffentlicht: (2025)
von: Lai, Peichao, et al.
Veröffentlicht: (2025)
"I understand why I got this grade": Automatic Short Answer Grading with Feedback
von: Aggarwal, Dishank, et al.
Veröffentlicht: (2024)
von: Aggarwal, Dishank, et al.
Veröffentlicht: (2024)
Quality-Conditioned Agreement in Automated Short Answer Scoring: Mid-Range Degradation and the Impact of Task-Specific Adaptation
von: Schleifer, Abigail Victoria Gurin, et al.
Veröffentlicht: (2026)
von: Schleifer, Abigail Victoria Gurin, et al.
Veröffentlicht: (2026)
From Flat to Structural: Enhancing Automated Short Answer Grading with GraphRAG
von: Chu, Yucheng, et al.
Veröffentlicht: (2026)
von: Chu, Yucheng, et al.
Veröffentlicht: (2026)
Automated Essay Scoring Incorporating Annotations from Automated Feedback Systems
von: Ormerod, Christopher
Veröffentlicht: (2025)
von: Ormerod, Christopher
Veröffentlicht: (2025)
Beyond-RAG: Question Identification and Answer Generation in Real-Time Conversations
von: Agrawal, Garima, et al.
Veröffentlicht: (2024)
von: Agrawal, Garima, et al.
Veröffentlicht: (2024)
Automatic Essay Multi-dimensional Scoring with Fine-tuning and Multiple Regression
von: Sun, Kun, et al.
Veröffentlicht: (2024)
von: Sun, Kun, et al.
Veröffentlicht: (2024)
Beyond Pointwise Scores: Decomposed Criteria-Based Evaluation of LLM Responses
von: Yu, Fangyi, et al.
Veröffentlicht: (2025)
von: Yu, Fangyi, et al.
Veröffentlicht: (2025)
Improving Score Reliability of Multiple Choice Benchmarks with Consistency Evaluation and Altered Answer Choices
von: Cavalin, Paulo, et al.
Veröffentlicht: (2025)
von: Cavalin, Paulo, et al.
Veröffentlicht: (2025)
Knowledge Distillation of LLM for Automatic Scoring of Science Education Assessments
von: Latif, Ehsan, et al.
Veröffentlicht: (2023)
von: Latif, Ehsan, et al.
Veröffentlicht: (2023)
Applying Large Language Models and Chain-of-Thought for Automatic Scoring
von: Lee, Gyeong-Geon, et al.
Veröffentlicht: (2023)
von: Lee, Gyeong-Geon, et al.
Veröffentlicht: (2023)
Improve LLM-based Automatic Essay Scoring with Linguistic Features
von: Hou, Zhaoyi Joey, et al.
Veröffentlicht: (2025)
von: Hou, Zhaoyi Joey, et al.
Veröffentlicht: (2025)
Artificial Intelligence Bias on English Language Learners in Automatic Scoring
von: Guo, Shuchen, et al.
Veröffentlicht: (2025)
von: Guo, Shuchen, et al.
Veröffentlicht: (2025)
Modular RAG: Transforming RAG Systems into LEGO-like Reconfigurable Frameworks
von: Gao, Yunfan, et al.
Veröffentlicht: (2024)
von: Gao, Yunfan, et al.
Veröffentlicht: (2024)
DiffScore: Text Evaluation Beyond Autoregressive Likelihood
von: Lai, Wen, et al.
Veröffentlicht: (2026)
von: Lai, Wen, et al.
Veröffentlicht: (2026)
Can Large Language Models Automatically Score Proficiency of Written Essays?
von: Mansour, Watheq, et al.
Veröffentlicht: (2024)
von: Mansour, Watheq, et al.
Veröffentlicht: (2024)
Using GPT-4 to Augment Unbalanced Data for Automatic Scoring
von: Fang, Luyang, et al.
Veröffentlicht: (2023)
von: Fang, Luyang, et al.
Veröffentlicht: (2023)
Fine-tuning ChatGPT for Automatic Scoring of Written Scientific Explanations in Chinese
von: Yang, Jie, et al.
Veröffentlicht: (2025)
von: Yang, Jie, et al.
Veröffentlicht: (2025)
Rank-Then-Score: Enhancing Large Language Models for Automated Essay Scoring
von: Cai, Yida, et al.
Veröffentlicht: (2025)
von: Cai, Yida, et al.
Veröffentlicht: (2025)
Autoregressive Score Generation for Multi-trait Essay Scoring
von: Do, Heejin, et al.
Veröffentlicht: (2024)
von: Do, Heejin, et al.
Veröffentlicht: (2024)
Beyond Bias Scores: Unmasking Vacuous Neutrality in Small Language Models
von: Manduru, Sumanth, et al.
Veröffentlicht: (2025)
von: Manduru, Sumanth, et al.
Veröffentlicht: (2025)
Summarization Metrics for Spanish and Basque: Do Automatic Scores and LLM-Judges Correlate with Humans?
von: Barnes, Jeremy, et al.
Veröffentlicht: (2025)
von: Barnes, Jeremy, et al.
Veröffentlicht: (2025)
An Open Multilingual System for Scoring Readability of Wikipedia
von: Trokhymovych, Mykola, et al.
Veröffentlicht: (2024)
von: Trokhymovych, Mykola, et al.
Veröffentlicht: (2024)
Confident RAG: Enhancing the Performance of LLMs for Mathematics Question Answering through Multi-Embedding and Confidence Scoring
von: Chen, Shiting, et al.
Veröffentlicht: (2025)
von: Chen, Shiting, et al.
Veröffentlicht: (2025)
Hallucination-Free Automatic Question & Answer Generation for Intuitive Learning
von: Wang, Nicholas X., et al.
Veröffentlicht: (2026)
von: Wang, Nicholas X., et al.
Veröffentlicht: (2026)
Beyond Agreement: Scoring Panel-Surfaced Biomedical Entity Candidates for Curator Triage
von: Cao, Shuheng, et al.
Veröffentlicht: (2026)
von: Cao, Shuheng, et al.
Veröffentlicht: (2026)
Pair2Score: Pairwise-to-Absolute Transfer for LLM-Based Essay Scoring
von: Hallaç, İbrahim Rıza, et al.
Veröffentlicht: (2026)
von: Hallaç, İbrahim Rıza, et al.
Veröffentlicht: (2026)
AdaGReS:Adaptive Greedy Context Selection via Redundancy-Aware Scoring for Token-Budgeted RAG
von: Peng, Chao, et al.
Veröffentlicht: (2025)
von: Peng, Chao, et al.
Veröffentlicht: (2025)
Towards LLM-based Autograding for Short Textual Answers
von: Schneider, Johannes, et al.
Veröffentlicht: (2023)
von: Schneider, Johannes, et al.
Veröffentlicht: (2023)
StackRAG Agent: Improving Developer Answers with Retrieval-Augmented Generation
von: Abrahamyan, Davit, et al.
Veröffentlicht: (2024)
von: Abrahamyan, Davit, et al.
Veröffentlicht: (2024)
Teach-to-Reason with Scoring: Self-Explainable Rationale-Driven Multi-Trait Essay Scoring
von: Do, Heejin, et al.
Veröffentlicht: (2025)
von: Do, Heejin, et al.
Veröffentlicht: (2025)
Transformer-based Joint Modelling for Automatic Essay Scoring and Off-Topic Detection
von: Das, Sourya Dipta, et al.
Veröffentlicht: (2024)
von: Das, Sourya Dipta, et al.
Veröffentlicht: (2024)
Autoregressive Multi-trait Essay Scoring via Reinforcement Learning with Scoring-aware Multiple Rewards
von: Do, Heejin, et al.
Veröffentlicht: (2024)
von: Do, Heejin, et al.
Veröffentlicht: (2024)
Directed Graph-alignment Approach for Identification of Gaps in Short Answers
von: Sahu, Archana, et al.
Veröffentlicht: (2025)
von: Sahu, Archana, et al.
Veröffentlicht: (2025)
Enhancing Essay Scoring with Adversarial Weights Perturbation and Metric-specific AttentionPooling
von: Huang, Jiaxin, et al.
Veröffentlicht: (2024)
von: Huang, Jiaxin, et al.
Veröffentlicht: (2024)
From Scores to Steps: Diagnosing and Improving LLM Performance in Evidence-Based Medical Calculations
von: Wang, Benlu, et al.
Veröffentlicht: (2025)
von: Wang, Benlu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Developing a Tutoring Dialog Dataset to Optimize LLMs for Educational Use
von: Fateen, Menna, et al.
Veröffentlicht: (2024) -
One Stone, Four Birds: A Comprehensive Solution for QA System Using Supervised Contrastive Learning
von: Wang, Bo, et al.
Veröffentlicht: (2024) -
Automatic Essay Scoring and Feedback Generation in Basque Language Learning
von: Azurmendi, Ekhi, et al.
Veröffentlicht: (2025) -
Generative Language Models with Retrieval Augmented Generation for Automated Short Answer Scoring
von: Wang, Zifan, et al.
Veröffentlicht: (2024) -
SAS-Bench: A Fine-Grained Benchmark for Evaluating Short Answer Scoring with Large Language Models
von: Lai, Peichao, et al.
Veröffentlicht: (2025)