OpenLearnLM Benchmark: A Unified Framework for Evaluating Knowledge, Skill, and Attitude in Educational Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Lee, Unggi, Lee, Sookbun, Choi, Heungsoo, Lee, Jinseo, Park, Haeun, Jeon, Younghoon, Cho, Sungmin, Kang, Minju, Koh, Junbo, Bae, Jiyeong, Nam, Minwoo, Eun, Juyeon, Jung, Yeonji, Jeong, Yeil |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Rewarding How Models Think Pedagogically: Integrating Pedagogical Reasoning and Thinking Rewards for LLMs in Education
di: Lee, Unggi, et al.
Pubblicazione: (2026)
di: Lee, Unggi, et al.
Pubblicazione: (2026)
Exploring the Role of Automated Feedback in Programming Education: A Systematic Literature Review
di: Jung, Yeonji, et al.
Pubblicazione: (2026)
di: Jung, Yeonji, et al.
Pubblicazione: (2026)
Pedagogy-R1: Pedagogically-Aligned Reasoning Model with Balanced Educational Benchmark
di: Lee, Unggi, et al.
Pubblicazione: (2025)
di: Lee, Unggi, et al.
Pubblicazione: (2025)
ISD-Agent-Bench: A Comprehensive Benchmark for Evaluating LLM-based Instructional Design Agents
di: Jeon, YoungHoon, et al.
Pubblicazione: (2026)
di: Jeon, YoungHoon, et al.
Pubblicazione: (2026)
Reinforcement Learning for Special Education: Aligning LLM Tutors to Diverse Learners through Disability-Adaptive Training
di: Lee, Unggi, et al.
Pubblicazione: (2026)
di: Lee, Unggi, et al.
Pubblicazione: (2026)
TeachObs: A Human-Validated Benchmark for Multimodal Teaching Observation and Model Evaluation
di: Jeong, Yeil, et al.
Pubblicazione: (2026)
di: Jeong, Yeil, et al.
Pubblicazione: (2026)
LLMs Are Already Good Tutors: Training-Free Prompt Optimization for Pedagogical Math Tutoring
di: Lee, Unggi, et al.
Pubblicazione: (2026)
di: Lee, Unggi, et al.
Pubblicazione: (2026)
ES-KT-24: A Multimodal Knowledge Tracing Benchmark Dataset with Educational Game Playing Video and Synthetic Text Generation
di: Kim, Dohee, et al.
Pubblicazione: (2024)
di: Kim, Dohee, et al.
Pubblicazione: (2024)
From Prediction to Application: Language Model-based Code Knowledge Tracing with Domain Adaptive Pre-Training and Automatic Feedback System with Pedagogical Prompting for Comprehensive Programming Education
di: Lee, Unggi, et al.
Pubblicazione: (2024)
di: Lee, Unggi, et al.
Pubblicazione: (2024)
I See You: Teacher Analytics with GPT-4 Vision-Powered Observational Assessment
di: Lee, Unggi, et al.
Pubblicazione: (2024)
di: Lee, Unggi, et al.
Pubblicazione: (2024)
Language Model Can Do Knowledge Tracing: Simple but Effective Method to Integrate Language Model and Knowledge Tracing Task
di: Lee, Unggi, et al.
Pubblicazione: (2024)
di: Lee, Unggi, et al.
Pubblicazione: (2024)
How to Align Large Language Models for Teaching English? Designing and Developing LLM based-Chatbot for Teaching English Conversation in EFL, Findings and Limitations
di: Park, Jaekwon, et al.
Pubblicazione: (2024)
di: Park, Jaekwon, et al.
Pubblicazione: (2024)
Llama-Polya: Instruction Tuning for Large Language Model based on Polya's Problem-solving
di: Lee, Unggi, et al.
Pubblicazione: (2026)
di: Lee, Unggi, et al.
Pubblicazione: (2026)
Are Video Models Zero-Shot Learners and Reasoners in Education? EduVideoBench, A Knowledge-Skills-Attitude Benchmark for Educational Video Generation
di: Lee, Unggi, et al.
Pubblicazione: (2026)
di: Lee, Unggi, et al.
Pubblicazione: (2026)
Return Prediction for Mean-Variance Portfolio Selection: How Decision-Focused Learning Shapes Forecasting Models
di: Lee, Junhyeong, et al.
Pubblicazione: (2024)
di: Lee, Junhyeong, et al.
Pubblicazione: (2024)
Color Information-Based Automated Mask Generation for Detecting Underwater Atypical Glare Areas
di: Jeon, Mingyu, et al.
Pubblicazione: (2025)
di: Jeon, Mingyu, et al.
Pubblicazione: (2025)
How Can Video Generative AI Transform K-12 Education? Examining Teachers' Perspectives through TPACK and TAM
di: Lee, Unggi, et al.
Pubblicazione: (2025)
di: Lee, Unggi, et al.
Pubblicazione: (2025)
Class size and school gender composition
di: Jiyeong Lee
Pubblicazione: (2025)
di: Jiyeong Lee
Pubblicazione: (2025)
A Cholesky decomposition-based asset selection heuristic for sparse tangent portfolio optimization
di: Bae, Hyunglip, et al.
Pubblicazione: (2025)
di: Bae, Hyunglip, et al.
Pubblicazione: (2025)
Prediction Loss Guided Decision-Focused Learning
di: Jeon, Haeun, et al.
Pubblicazione: (2025)
di: Jeon, Haeun, et al.
Pubblicazione: (2025)
Unilateral biportal endoscopic partial cervical laminectomy and facetectomy: An ex vivo study and case report
di: Hojung Bae, et al.
Pubblicazione: (2026)
di: Hojung Bae, et al.
Pubblicazione: (2026)
LLMs Have Made Failure Worth Publishing
di: Lee, Sungmin
Pubblicazione: (2026)
di: Lee, Sungmin
Pubblicazione: (2026)
LLM-based Question-Answer Framework for Sensor-driven HVAC System Interaction
di: Lee, Sungmin, et al.
Pubblicazione: (2025)
di: Lee, Sungmin, et al.
Pubblicazione: (2025)
The Effect of Age Identity on Depression in Middle-Aged and Older Adults: The Mediating Role of Attitudes Toward Aging and the Moderated Mediation Effect of Subjective Health Status
di: Lee, Ki-Nam, et al.
Pubblicazione: (2025)
di: Lee, Ki-Nam, et al.
Pubblicazione: (2025)
What Drives Bricolage? The Effects of Learning Orientation, Resource Constraints and Environmental Turbulence
di: Juyeon Lee, et al.
Pubblicazione: (2026)
di: Juyeon Lee, et al.
Pubblicazione: (2026)
Tailoring B‐Doped Co(OH) x Nanosheets for Quasi‐Homogeneous Catalytic Hydrogenation of 4‐Nitrophenol
di: Sunglun Kwon, et al.
Pubblicazione: (2026)
di: Sunglun Kwon, et al.
Pubblicazione: (2026)
Verifiable Dropout: Turning Randomness into a Verifiable Claim
di: Lee, Kichang, et al.
Pubblicazione: (2025)
di: Lee, Kichang, et al.
Pubblicazione: (2025)
Spatial Discretization for Fine-Grain Zone Checks with STARKs
di: Lee, Sungmin, et al.
Pubblicazione: (2025)
di: Lee, Sungmin, et al.
Pubblicazione: (2025)
Can LLMs Recognize Toxicity? A Structured Investigation Framework and Toxicity Metric
di: Koh, Hyukhun, et al.
Pubblicazione: (2024)
di: Koh, Hyukhun, et al.
Pubblicazione: (2024)
Fine-grained Gender Control in Machine Translation with Large Language Models
di: Lee, Minwoo, et al.
Pubblicazione: (2024)
di: Lee, Minwoo, et al.
Pubblicazione: (2024)
Unraveling Human Capital Complexity: Economic Complexity Analysis of Occupations and Skills
di: Lee, Soohyoung, et al.
Pubblicazione: (2025)
di: Lee, Soohyoung, et al.
Pubblicazione: (2025)
HCN and HNC in the Disk of an Outbursting Young Star, V883 Ori
di: Lee, Seonjae, et al.
Pubblicazione: (2024)
di: Lee, Seonjae, et al.
Pubblicazione: (2024)
Characterization of synoptic environment for mesoscale convective systems over South Korea using ERA5 reanalysis data
di: Jeong‐Eun Lee, et al.
Pubblicazione: (2026)
di: Jeong‐Eun Lee, et al.
Pubblicazione: (2026)
A Review on Soft Ionic Touch Point Sensors
di: Gibeom Lee, et al.
Pubblicazione: (2024)
di: Gibeom Lee, et al.
Pubblicazione: (2024)
Contrast-CAT: Contrasting Activations for Enhanced Interpretability in Transformer-based Text Classifiers
di: Han, Sungmin, et al.
Pubblicazione: (2025)
di: Han, Sungmin, et al.
Pubblicazione: (2025)
Enhancement of Bone Tissue Regeneration with Multi‐Functional Nanoparticles by Coordination of Immune, Osteogenic, and Angiogenic Responses (Adv. Healthcare Mater. 5/2025)
di: Hyewoo Jeong, et al.
Pubblicazione: (2025)
di: Hyewoo Jeong, et al.
Pubblicazione: (2025)
Group-wise Scaling and Orthogonal Decomposition for Domain-Invariant Feature Extraction in Face Anti-Spoofing
di: Jung, Seungjin, et al.
Pubblicazione: (2025)
di: Jung, Seungjin, et al.
Pubblicazione: (2025)
Properness and finiteness of totally geodesic submanifolds in the convex core
di: Lee, Minju, et al.
Pubblicazione: (2025)
di: Lee, Minju, et al.
Pubblicazione: (2025)
Discrete subgroups with finite Bowen-Margulis-Sullivan measure in higher rank
di: Fraczyk, Mikolaj, et al.
Pubblicazione: (2023)
di: Fraczyk, Mikolaj, et al.
Pubblicazione: (2023)
Orbit closures of unipotent flows for hyperbolic manifolds with Fuchsian ends
di: Lee, Minju, et al.
Pubblicazione: (2019)
di: Lee, Minju, et al.
Pubblicazione: (2019)
Documenti analoghi
-
Rewarding How Models Think Pedagogically: Integrating Pedagogical Reasoning and Thinking Rewards for LLMs in Education
di: Lee, Unggi, et al.
Pubblicazione: (2026) -
Exploring the Role of Automated Feedback in Programming Education: A Systematic Literature Review
di: Jung, Yeonji, et al.
Pubblicazione: (2026) -
Pedagogy-R1: Pedagogically-Aligned Reasoning Model with Balanced Educational Benchmark
di: Lee, Unggi, et al.
Pubblicazione: (2025) -
ISD-Agent-Bench: A Comprehensive Benchmark for Evaluating LLM-based Instructional Design Agents
di: Jeon, YoungHoon, et al.
Pubblicazione: (2026) -
Reinforcement Learning for Special Education: Aligning LLM Tutors to Diverse Learners through Disability-Adaptive Training
di: Lee, Unggi, et al.
Pubblicazione: (2026)