SID: Benchmarking Guided Instruction Capabilities in STEM Education with a Socratic Interdisciplinary Dialogues Dataset
Fuente:
arXiv
Saved in:
| Main Authors: | Jiang, Mei, Yue, Houping, Li, Bingdong, Hao, Hao, Qian, Ying, Jiang, Bo, Zhou, Aimin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
EduResearchBench: A Hierarchical Atomic Task Decomposition Benchmark for Full-Lifecycle Educational Research
by: Yue, Houping, et al.
Published: (2026)
by: Yue, Houping, et al.
Published: (2026)
Evolutionary Reinforcement Learning based AI tutor for Socratic Interdisciplinary Instruction
by: Jiang, Mei, et al.
Published: (2025)
by: Jiang, Mei, et al.
Published: (2025)
IB-GRPO: Aligning LLM-based Learning Path Recommendation with Educational Objectives via Indicator-Based Group Relative Policy Optimization
by: Wang, Shuai, et al.
Published: (2026)
by: Wang, Shuai, et al.
Published: (2026)
Surrogate-Assisted Evolutionary Reinforcement Learning Based on Autoencoder and Hyperbolic Neural Network
by: Li, Bingdong, et al.
Published: (2025)
by: Li, Bingdong, et al.
Published: (2025)
ORCDF: An Oversmoothing-Resistant Cognitive Diagnosis Framework for Student Learning in Online Education Systems
by: Qian, Hong, et al.
Published: (2024)
by: Qian, Hong, et al.
Published: (2024)
How Real Is AI Tutoring? Comparing Simulated and Human Dialogues in One-on-One Instruction
by: Li, Ruijia, et al.
Published: (2025)
by: Li, Ruijia, et al.
Published: (2025)
CSVQA: A Chinese Multimodal Benchmark for Evaluating STEM Reasoning Capabilities of VLMs
by: Jian, Ai, et al.
Published: (2025)
by: Jian, Ai, et al.
Published: (2025)
Expensive Multi-Objective Bayesian Optimization Based on Diffusion Models
by: Li, Bingdong, et al.
Published: (2024)
by: Li, Bingdong, et al.
Published: (2024)
Context-aware Diversity Enhancement for Neural Multi-Objective Combinatorial Optimization
by: Lu, Yongfan, et al.
Published: (2024)
by: Lu, Yongfan, et al.
Published: (2024)
Scaling Laws for Educational AI Agents
by: Wu, Mengsong, et al.
Published: (2026)
by: Wu, Mengsong, et al.
Published: (2026)
See and Remember: A Multimodal Agent for Web Traversal
by: Wang, Xinjun, et al.
Published: (2026)
by: Wang, Xinjun, et al.
Published: (2026)
Closing the Expression Gap in LLM Instructions via Socratic Questioning
by: Sun, Jianwen, et al.
Published: (2025)
by: Sun, Jianwen, et al.
Published: (2025)
Inductive Cognitive Diagnosis for Fast Student Learning in Web-Based Online Intelligent Education Systems
by: Liu, Shuo, et al.
Published: (2024)
by: Liu, Shuo, et al.
Published: (2024)
EA4LLM: A Gradient-Free Approach to Large Language Model Optimization via Evolutionary Algorithms
by: Liu, WenTao, et al.
Published: (2025)
by: Liu, WenTao, et al.
Published: (2025)
A Study on Educational Data Analysis and Personalized Feedback Report Generation Based on Tags and ChatGPT
by: Zhou, Yizhou, et al.
Published: (2025)
by: Zhou, Yizhou, et al.
Published: (2025)
LLM-KT: Aligning Large Language Models with Knowledge Tracing using a Plug-and-Play Instruction
by: Wang, Ziwei, et al.
Published: (2025)
by: Wang, Ziwei, et al.
Published: (2025)
Biology-Instructions: A Dataset and Benchmark for Multi-Omics Sequence Understanding Capability of Large Language Models
by: He, Haonan, et al.
Published: (2024)
by: He, Haonan, et al.
Published: (2024)
Data Selection for Multi-turn Dialogue Instruction Tuning
by: Li, Bo, et al.
Published: (2026)
by: Li, Bo, et al.
Published: (2026)
Unlocking Proactivity in Task-Oriented Dialogue
by: Zhang, Hongbin, et al.
Published: (2026)
by: Zhang, Hongbin, et al.
Published: (2026)
Evolutionary Retrosynthetic Route Planning
by: Zhang, Yan, et al.
Published: (2023)
by: Zhang, Yan, et al.
Published: (2023)
Flow2Code: Evaluating Large Language Models for Flowchart-based Code Generation Capability
by: He, Mengliang, et al.
Published: (2025)
by: He, Mengliang, et al.
Published: (2025)
STEM: Efficient Relative Capability Evaluation of LLMs through Structured Transition Samples
by: Hu, Haiquan, et al.
Published: (2025)
by: Hu, Haiquan, et al.
Published: (2025)
Discerning minds or generic tutors? Evaluating instructional guidance capabilities in Socratic LLMs
by: Liu, Ying, et al.
Published: (2025)
by: Liu, Ying, et al.
Published: (2025)
ASL STEM Wiki: Dataset and Benchmark for Interpreting STEM Articles
by: Yin, Kayo, et al.
Published: (2024)
by: Yin, Kayo, et al.
Published: (2024)
Prompt-SID: Learning Structural Representation Prompt via Latent Diffusion for Single-Image Denoising
by: Li, Huaqiu, et al.
Published: (2025)
by: Li, Huaqiu, et al.
Published: (2025)
Agentic Workflow for Education: Concepts and Applications
by: Jiang, Yuan-Hao, et al.
Published: (2025)
by: Jiang, Yuan-Hao, et al.
Published: (2025)
Symbolic Cognitive Diagnosis via Hybrid Optimization for Intelligent Education Systems
by: Shen, Junhao, et al.
Published: (2023)
by: Shen, Junhao, et al.
Published: (2023)
SSR: Socratic Self-Refine for Large Language Model Reasoning
by: Shi, Haizhou, et al.
Published: (2025)
by: Shi, Haizhou, et al.
Published: (2025)
A Dialogue-Based Framework for Correcting Multimodal Errors in AI-Assisted STEM Education
by: Syal, Akshay, et al.
Published: (2026)
by: Syal, Akshay, et al.
Published: (2026)
Uni-Retrieval: A Multi-Style Retrieval Framework for STEM's Education
by: Jia, Yanhao, et al.
Published: (2025)
by: Jia, Yanhao, et al.
Published: (2025)
A First Look at Kolmogorov-Arnold Networks in Surrogate-assisted Evolutionary Algorithms
by: Hao, Hao, et al.
Published: (2024)
by: Hao, Hao, et al.
Published: (2024)
Relation Reasoning with LLMs in Expensive Optimization
by: Lu, Ye, et al.
Published: (2026)
by: Lu, Ye, et al.
Published: (2026)
DialogueReason: Rule-Based RL Sparks Dialogue Reasoning in LLMs
by: Shu, Yubo, et al.
Published: (2025)
by: Shu, Yubo, et al.
Published: (2025)
PandaLM: An Automatic Evaluation Benchmark for LLM Instruction Tuning Optimization
by: Wang, Yidong, et al.
Published: (2023)
by: Wang, Yidong, et al.
Published: (2023)
DialogueForge: LLM Simulation of Human-Chatbot Dialogue
by: Zhu, Ruizhe, et al.
Published: (2025)
by: Zhu, Ruizhe, et al.
Published: (2025)
AgentSchool: An LLM-Powered Multi-Agent Simulation for Education
by: Ye, Yulei, et al.
Published: (2026)
by: Ye, Yulei, et al.
Published: (2026)
Enhancing Explainability of Knowledge Learning Paths: Causal Knowledge Networks
by: Wei, Yuang, et al.
Published: (2024)
by: Wei, Yuang, et al.
Published: (2024)
Refine Large Language Model Fine-tuning via Instruction Vector
by: Jiang, Gangwei, et al.
Published: (2024)
by: Jiang, Gangwei, et al.
Published: (2024)
Mutual Reinforcement of LLM Dialogue Synthesis and Summarization Capabilities for Few-Shot Dialogue Summarization
by: Lu, Yen-Ju, et al.
Published: (2025)
by: Lu, Yen-Ju, et al.
Published: (2025)
LeCoDe: A Benchmark Dataset for Interactive Legal Consultation Dialogue Evaluation
by: Yuan, Weikang, et al.
Published: (2025)
by: Yuan, Weikang, et al.
Published: (2025)
Similar Items
-
EduResearchBench: A Hierarchical Atomic Task Decomposition Benchmark for Full-Lifecycle Educational Research
by: Yue, Houping, et al.
Published: (2026) -
Evolutionary Reinforcement Learning based AI tutor for Socratic Interdisciplinary Instruction
by: Jiang, Mei, et al.
Published: (2025) -
IB-GRPO: Aligning LLM-based Learning Path Recommendation with Educational Objectives via Indicator-Based Group Relative Policy Optimization
by: Wang, Shuai, et al.
Published: (2026) -
Surrogate-Assisted Evolutionary Reinforcement Learning Based on Autoencoder and Hyperbolic Neural Network
by: Li, Bingdong, et al.
Published: (2025) -
ORCDF: An Oversmoothing-Resistant Cognitive Diagnosis Framework for Student Learning in Online Education Systems
by: Qian, Hong, et al.
Published: (2024)