Knowledge Distillation of LLM for Automatic Scoring of Science Education Assessments
Fuente:
arXiv
Saved in:
| Main Authors: | Latif, Ehsan, Fang, Luyang, Ma, Ping, Zhai, Xiaoming |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fine-tuning ChatGPT for Automatic Scoring of Written Scientific Explanations in Chinese
by: Yang, Jie, et al.
Published: (2025)
by: Yang, Jie, et al.
Published: (2025)
Using GPT-4 to Augment Unbalanced Data for Automatic Scoring
by: Fang, Luyang, et al.
Published: (2023)
by: Fang, Luyang, et al.
Published: (2023)
Applying Large Language Models and Chain-of-Thought for Automatic Scoring
by: Lee, Gyeong-Geon, et al.
Published: (2023)
by: Lee, Gyeong-Geon, et al.
Published: (2023)
G-SciEdBERT: A Contextualized LLM for Science Assessment Tasks in German
by: Latif, Ehsan, et al.
Published: (2024)
by: Latif, Ehsan, et al.
Published: (2024)
Gemini Pro Defeated by GPT-4V: Evidence from Education
by: Lee, Gyeong-Geon, et al.
Published: (2023)
by: Lee, Gyeong-Geon, et al.
Published: (2023)
Efficient Multi-Task Inferencing with a Shared Backbone and Lightweight Task-Specific Adapters for Automatic Scoring
by: Latif, Ehsan, et al.
Published: (2024)
by: Latif, Ehsan, et al.
Published: (2024)
Privacy-Preserved Automated Scoring using Federated Learning for Educational Research
by: Latif, Ehsan, et al.
Published: (2025)
by: Latif, Ehsan, et al.
Published: (2025)
Artificial Intelligence Bias on English Language Learners in Automatic Scoring
by: Guo, Shuchen, et al.
Published: (2025)
by: Guo, Shuchen, et al.
Published: (2025)
Efficient Multi-Task Inferencing: Model Merging with Gromov-Wasserstein Feature Alignment
by: Fang, Luyang, et al.
Published: (2025)
by: Fang, Luyang, et al.
Published: (2025)
SketchMind: A Multi-Agent Cognitive Framework for Assessing Student-Drawn Scientific Sketches
by: Latif, Ehsan, et al.
Published: (2025)
by: Latif, Ehsan, et al.
Published: (2025)
Unveiling Scoring Processes: Dissecting the Differences between LLMs and Human Graders in Automatic Scoring
by: Wu, Xuansheng, et al.
Published: (2024)
by: Wu, Xuansheng, et al.
Published: (2024)
Improve LLM-based Automatic Essay Scoring with Linguistic Features
by: Hou, Zhaoyi Joey, et al.
Published: (2025)
by: Hou, Zhaoyi Joey, et al.
Published: (2025)
LLM-Oriented Token-Adaptive Knowledge Distillation
by: Xie, Xurong, et al.
Published: (2025)
by: Xie, Xurong, et al.
Published: (2025)
Summarization Metrics for Spanish and Basque: Do Automatic Scores and LLM-Judges Correlate with Humans?
by: Barnes, Jeremy, et al.
Published: (2025)
by: Barnes, Jeremy, et al.
Published: (2025)
Automatic Question & Answer Generation Using Generative Large Language Model (LLM)
by: Ehsan, Md. Alvee, et al.
Published: (2025)
by: Ehsan, Md. Alvee, et al.
Published: (2025)
Using Generative AI and Multi-Agents to Provide Automatic Feedback
by: Guo, Shuchen, et al.
Published: (2024)
by: Guo, Shuchen, et al.
Published: (2024)
AutoSCORE: Enhancing Automated Scoring with Multi-Agent Large Language Models via Structured Component Recognition
by: Wang, Yun, et al.
Published: (2025)
by: Wang, Yun, et al.
Published: (2025)
BRIDGE the Gap: Mitigating Bias Amplification in Automated Scoring of English Language Learners via Inter-group Data Augmentation
by: Wang, Yun, et al.
Published: (2026)
by: Wang, Yun, et al.
Published: (2026)
DistillGuard: Evaluating Defenses Against LLM Knowledge Distillation
by: Jiang, Bo
Published: (2026)
by: Jiang, Bo
Published: (2026)
From Deferral to Learning: Online In-Context Knowledge Distillation for LLM Cascades
by: Wu, Yu, et al.
Published: (2025)
by: Wu, Yu, et al.
Published: (2025)
How Reliable Are Automatic Evaluation Methods for Instruction-Tuned LLMs?
by: Doostmohammadi, Ehsan, et al.
Published: (2024)
by: Doostmohammadi, Ehsan, et al.
Published: (2024)
Next Token Perception Score: Analytical Assessment of your LLM Perception Skills
by: Cheng, Yu-Ang, et al.
Published: (2025)
by: Cheng, Yu-Ang, et al.
Published: (2025)
Towards Generalization of Block Attention via Automatic Segmentation and Block Distillation
by: Li, Shuaiyi, et al.
Published: (2026)
by: Li, Shuaiyi, et al.
Published: (2026)
Automatic Generation of Inference Making Questions for Reading Comprehension Assessments
by: Ma, Wanjing Anya, et al.
Published: (2025)
by: Ma, Wanjing Anya, et al.
Published: (2025)
Reliable Reasoning Path: Distilling Effective Guidance for LLM Reasoning with Knowledge Graphs
by: Xiao, Yilin, et al.
Published: (2025)
by: Xiao, Yilin, et al.
Published: (2025)
Assessing the potential of LLM-assisted annotation for corpus-based pragmatics and discourse analysis: The case of apology
by: Yu, Danni, et al.
Published: (2023)
by: Yu, Danni, et al.
Published: (2023)
Distilling Closed-Source LLM's Knowledge for Locally Stable and Economic Biomedical Entity Linking
by: Ai, Yihao, et al.
Published: (2025)
by: Ai, Yihao, et al.
Published: (2025)
Beyond Scores: A Modular RAG-Based System for Automatic Short Answer Scoring with Feedback
by: Fateen, Menna, et al.
Published: (2024)
by: Fateen, Menna, et al.
Published: (2024)
AI and Machine Learning for Next Generation Science Assessments
by: Zhai, Xiaoming
Published: (2024)
by: Zhai, Xiaoming
Published: (2024)
Flow-Modulated Scoring for Semantic-Aware Knowledge Graph Completion
by: Li, Siyuan, et al.
Published: (2025)
by: Li, Siyuan, et al.
Published: (2025)
The Landscape of AI in Science Education: What is Changing and How to Respond
by: Zhai, Xiaoming, et al.
Published: (2026)
by: Zhai, Xiaoming, et al.
Published: (2026)
Automatic Essay Scoring and Feedback Generation in Basque Language Learning
by: Azurmendi, Ekhi, et al.
Published: (2025)
by: Azurmendi, Ekhi, et al.
Published: (2025)
Pedagogically-Inspired Data Synthesis for Language Model Knowledge Distillation
by: He, Bowei, et al.
Published: (2026)
by: He, Bowei, et al.
Published: (2026)
Silicon Bureaucracy and AI Test-Oriented Education: Contamination Sensitivity and Score Confidence in LLM Benchmarks
by: Song, Yiliang, et al.
Published: (2026)
by: Song, Yiliang, et al.
Published: (2026)
Implicit Bias in LLMs: A Survey
by: Lin, Xinru, et al.
Published: (2025)
by: Lin, Xinru, et al.
Published: (2025)
Generalizable and Efficient Automated Scoring with a Knowledge-Distilled Multi-Task Mixture-of-Experts
by: Fang, Luyang, et al.
Published: (2025)
by: Fang, Luyang, et al.
Published: (2025)
Automatic Essay Multi-dimensional Scoring with Fine-tuning and Multiple Regression
by: Sun, Kun, et al.
Published: (2024)
by: Sun, Kun, et al.
Published: (2024)
Can Large Language Models Automatically Score Proficiency of Written Essays?
by: Mansour, Watheq, et al.
Published: (2024)
by: Mansour, Watheq, et al.
Published: (2024)
Decoder-based Sense Knowledge Distillation
by: Wang, Qitong, et al.
Published: (2026)
by: Wang, Qitong, et al.
Published: (2026)
Self-Knowledge Distillation for Learning Ambiguity
by: Park, Hancheol, et al.
Published: (2024)
by: Park, Hancheol, et al.
Published: (2024)
Similar Items
-
Fine-tuning ChatGPT for Automatic Scoring of Written Scientific Explanations in Chinese
by: Yang, Jie, et al.
Published: (2025) -
Using GPT-4 to Augment Unbalanced Data for Automatic Scoring
by: Fang, Luyang, et al.
Published: (2023) -
Applying Large Language Models and Chain-of-Thought for Automatic Scoring
by: Lee, Gyeong-Geon, et al.
Published: (2023) -
G-SciEdBERT: A Contextualized LLM for Science Assessment Tasks in German
by: Latif, Ehsan, et al.
Published: (2024) -
Gemini Pro Defeated by GPT-4V: Evidence from Education
by: Lee, Gyeong-Geon, et al.
Published: (2023)