MalAlgoQA: Pedagogical Evaluation of Counterfactual Reasoning in Large Language Models and Implications for AI in Education
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Naiming, Sonkar, Shashank, Le, Myco, Baraniuk, Richard |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CLEAR-3K: Assessing Causal Explanatory Capabilities in Language Models
di: Liu, Naiming, et al.
Pubblicazione: (2025)
di: Liu, Naiming, et al.
Pubblicazione: (2025)
MetaCLASS: Metacognitive Coaching for Learning with Adaptive Self-regulation Support
di: Liu, Naiming, et al.
Pubblicazione: (2026)
di: Liu, Naiming, et al.
Pubblicazione: (2026)
Do LLMs Make Mistakes Like Students? Exploring Natural Alignment between Language Models and Human Error Patterns
di: Liu, Naiming, et al.
Pubblicazione: (2025)
di: Liu, Naiming, et al.
Pubblicazione: (2025)
Misconception Acquisition Dynamics in Large Language Models
di: Liu, Naiming, et al.
Pubblicazione: (2026)
di: Liu, Naiming, et al.
Pubblicazione: (2026)
LLM-based Cognitive Models of Students with Misconceptions
di: Sonkar, Shashank, et al.
Pubblicazione: (2024)
di: Sonkar, Shashank, et al.
Pubblicazione: (2024)
Student Data Paradox and Curious Case of Single Student-Tutor Model: Regressive Side Effects of Training LLMs for Personalized Learning
di: Sonkar, Shashank, et al.
Pubblicazione: (2024)
di: Sonkar, Shashank, et al.
Pubblicazione: (2024)
Pedagogical Alignment of Large Language Models
di: Sonkar, Shashank, et al.
Pubblicazione: (2024)
di: Sonkar, Shashank, et al.
Pubblicazione: (2024)
Atomic Learning Objectives Labeling: A High-Resolution Approach for Physics Education
di: Liu, Naiming, et al.
Pubblicazione: (2024)
di: Liu, Naiming, et al.
Pubblicazione: (2024)
MalruleLib: Large-Scale Executable Misconception Reasoning with Step Traces for Modeling Student Thinking in Mathematics
di: Chen, Xinghe, et al.
Pubblicazione: (2026)
di: Chen, Xinghe, et al.
Pubblicazione: (2026)
Marking: Visual Grading with Highlighting Errors and Annotating Missing Bits
di: Sonkar, Shashank, et al.
Pubblicazione: (2024)
di: Sonkar, Shashank, et al.
Pubblicazione: (2024)
FoundationalASSIST: An Educational Dataset for Foundational Knowledge Tracing and Pedagogical Grounding of LLMs
di: Worden, Eamon, et al.
Pubblicazione: (2026)
di: Worden, Eamon, et al.
Pubblicazione: (2026)
Scalable Generation and Validation of Isomorphic Physics Problems with GenAI
di: Liu, Naiming, et al.
Pubblicazione: (2026)
di: Liu, Naiming, et al.
Pubblicazione: (2026)
The Imitation Game for Educational AI
di: Sonkar, Shashank, et al.
Pubblicazione: (2025)
di: Sonkar, Shashank, et al.
Pubblicazione: (2025)
Many-Shot Regurgitation (MSR) Prompting
di: Sonkar, Shashank, et al.
Pubblicazione: (2024)
di: Sonkar, Shashank, et al.
Pubblicazione: (2024)
Training LLM-based Tutors to Improve Student Learning Outcomes in Dialogues
di: Scarlatos, Alexander, et al.
Pubblicazione: (2025)
di: Scarlatos, Alexander, et al.
Pubblicazione: (2025)
Circuit Complexity of Hierarchical Knowledge Tracing and Implications for Log-Precision Transformers
di: Liu, Naiming, et al.
Pubblicazione: (2026)
di: Liu, Naiming, et al.
Pubblicazione: (2026)
When Can We Trust LLM Graders? Calibrating Confidence for Automated Assessment
di: Ferrer, Robinson, et al.
Pubblicazione: (2026)
di: Ferrer, Robinson, et al.
Pubblicazione: (2026)
Simulating Students or Sycophantic Problem Solving? On Misconception Faithfulness of LLM Simulators
di: Do, Heejin, et al.
Pubblicazione: (2026)
di: Do, Heejin, et al.
Pubblicazione: (2026)
Synthetic Context Generation for Question Generation
di: Liu, Naiming, et al.
Pubblicazione: (2024)
di: Liu, Naiming, et al.
Pubblicazione: (2024)
Learning Context: A Unified Framework and Roadmap for Context-Aware AI in Education
di: Liu, Naiming, et al.
Pubblicazione: (2025)
di: Liu, Naiming, et al.
Pubblicazione: (2025)
MEDEQUALQA: Evaluating Biases in LLMs with Counterfactual Reasoning
di: Ghosh, Rajarshi, et al.
Pubblicazione: (2025)
di: Ghosh, Rajarshi, et al.
Pubblicazione: (2025)
GermanPartiesQA: Benchmarking Commercial Large Language Models and AI Companions for Political Alignment and Sycophancy
di: Batzner, Jan, et al.
Pubblicazione: (2024)
di: Batzner, Jan, et al.
Pubblicazione: (2024)
Large Language Models Approach Expert Pedagogical Quality in Math Tutoring but Differ in Instructional and Linguistic Profiles
di: Abdulsalam, Ramatu Oiza, et al.
Pubblicazione: (2025)
di: Abdulsalam, Ramatu Oiza, et al.
Pubblicazione: (2025)
Speciesism in AI: Evaluating Discrimination Against Animals in Large Language Models
di: Jotautaitė, Monika, et al.
Pubblicazione: (2025)
di: Jotautaitė, Monika, et al.
Pubblicazione: (2025)
Automated Long Answer Grading with RiceChem Dataset
di: Sonkar, Shashank, et al.
Pubblicazione: (2024)
di: Sonkar, Shashank, et al.
Pubblicazione: (2024)
SocialGaze: Improving the Integration of Human Social Norms in Large Language Models
di: Vijjini, Anvesh Rao, et al.
Pubblicazione: (2024)
di: Vijjini, Anvesh Rao, et al.
Pubblicazione: (2024)
Supervised Fine-Tuning LLMs to Behave as Pedagogical Agents in Programming Education
di: Ross, Emily, et al.
Pubblicazione: (2025)
di: Ross, Emily, et al.
Pubblicazione: (2025)
An Exploration of Higher Education Course Evaluation by Large Language Models
di: Yuan, Bo, et al.
Pubblicazione: (2024)
di: Yuan, Bo, et al.
Pubblicazione: (2024)
AI-VERDE: A Gateway for Egalitarian Access to Large Language Model-Based Resources For Educational Institutions
di: Mithun, Paul, et al.
Pubblicazione: (2025)
di: Mithun, Paul, et al.
Pubblicazione: (2025)
Assessing the Pedagogical Readiness of Large Language Models as AI Tutors in Low-Resource Contexts: A Case Study of Nepal's K-10 Curriculum
di: Acharya, Pratyush, et al.
Pubblicazione: (2026)
di: Acharya, Pratyush, et al.
Pubblicazione: (2026)
Question-Answering (QA) Model for a Personalized Learning Assistant for Arabic Language
di: Sammoudi, Mohammad, et al.
Pubblicazione: (2024)
di: Sammoudi, Mohammad, et al.
Pubblicazione: (2024)
Dr.Academy: A Benchmark for Evaluating Questioning Capability in Education for Large Language Models
di: Chen, Yuyan, et al.
Pubblicazione: (2024)
di: Chen, Yuyan, et al.
Pubblicazione: (2024)
ELMES: An Automated Framework for Evaluating Large Language Models in Educational Scenarios
di: Wei, Shou'ang, et al.
Pubblicazione: (2025)
di: Wei, Shou'ang, et al.
Pubblicazione: (2025)
Self-Blinding and Counterfactual Self-Simulation Mitigate Biases and Sycophancy in Large Language Models
di: Christian, Brian, et al.
Pubblicazione: (2026)
di: Christian, Brian, et al.
Pubblicazione: (2026)
AI Safety in Generative AI Large Language Models: A Survey
di: Chua, Jaymari, et al.
Pubblicazione: (2024)
di: Chua, Jaymari, et al.
Pubblicazione: (2024)
The World of Generative AI: Deepfakes and Large Language Models
di: Mitra, Alakananda, et al.
Pubblicazione: (2024)
di: Mitra, Alakananda, et al.
Pubblicazione: (2024)
Generalization in Healthcare AI: Evaluation of a Clinical Large Language Model
di: Rahman, Salman, et al.
Pubblicazione: (2024)
di: Rahman, Salman, et al.
Pubblicazione: (2024)
Evaluating Proactive Risk Awareness of Large Language Models
di: Luo, Xuan, et al.
Pubblicazione: (2026)
di: Luo, Xuan, et al.
Pubblicazione: (2026)
Extrinsic Evaluation of Cultural Competence in Large Language Models
di: Bhatt, Shaily, et al.
Pubblicazione: (2024)
di: Bhatt, Shaily, et al.
Pubblicazione: (2024)
Benchmarking Political Persuasion Risks Across Frontier Large Language Models
di: Chen, Zhongren, et al.
Pubblicazione: (2026)
di: Chen, Zhongren, et al.
Pubblicazione: (2026)
Documenti analoghi
-
CLEAR-3K: Assessing Causal Explanatory Capabilities in Language Models
di: Liu, Naiming, et al.
Pubblicazione: (2025) -
MetaCLASS: Metacognitive Coaching for Learning with Adaptive Self-regulation Support
di: Liu, Naiming, et al.
Pubblicazione: (2026) -
Do LLMs Make Mistakes Like Students? Exploring Natural Alignment between Language Models and Human Error Patterns
di: Liu, Naiming, et al.
Pubblicazione: (2025) -
Misconception Acquisition Dynamics in Large Language Models
di: Liu, Naiming, et al.
Pubblicazione: (2026) -
LLM-based Cognitive Models of Students with Misconceptions
di: Sonkar, Shashank, et al.
Pubblicazione: (2024)