Stepwise Verification and Remediation of Student Reasoning Errors with Large Language Model Tutors
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Daheim, Nico, Macina, Jakub, Kapur, Manu, Gurevych, Iryna, Sachan, Mrinmaya |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MathTutorBench: A Benchmark for Measuring Open-ended Pedagogical Capabilities of LLM Tutors
von: Macina, Jakub, et al.
Veröffentlicht: (2025)
von: Macina, Jakub, et al.
Veröffentlicht: (2025)
From Problem-Solving to Teaching Problem-Solving: Aligning LLMs with Pedagogy using Reinforcement Learning
von: Dinucu-Jianu, David, et al.
Veröffentlicht: (2025)
von: Dinucu-Jianu, David, et al.
Veröffentlicht: (2025)
Towards the Pedagogical Steering of Large Language Models for Tutoring: A Case Study with Modeling Productive Failure
von: Puech, Romain, et al.
Veröffentlicht: (2024)
von: Puech, Romain, et al.
Veröffentlicht: (2024)
Uncertainty-Aware Decoding with Minimum Bayes Risk
von: Daheim, Nico, et al.
Veröffentlicht: (2025)
von: Daheim, Nico, et al.
Veröffentlicht: (2025)
Model Merging by Uncertainty-Based Gradient Matching
von: Daheim, Nico, et al.
Veröffentlicht: (2023)
von: Daheim, Nico, et al.
Veröffentlicht: (2023)
Book2Dial: Generating Teacher-Student Interactions from Textbooks for Cost-Effective Development of Educational Chatbots
von: Wang, Junling, et al.
Veröffentlicht: (2024)
von: Wang, Junling, et al.
Veröffentlicht: (2024)
Improving Large Language Model Safety with Contrastive Representation Learning
von: Simko, Samuel, et al.
Veröffentlicht: (2025)
von: Simko, Samuel, et al.
Veröffentlicht: (2025)
SpaRC and SpaRP: Spatial Reasoning Characterization and Path Generation for Understanding Spatial Reasoning Capability of Large Language Models
von: Rizvi, Md Imbesat Hassan, et al.
Veröffentlicht: (2024)
von: Rizvi, Md Imbesat Hassan, et al.
Veröffentlicht: (2024)
How to Weight Multitask Finetuning? Fast Previews via Bayesian Model-Merging
von: Maldonado, Hugo Monzón, et al.
Veröffentlicht: (2024)
von: Maldonado, Hugo Monzón, et al.
Veröffentlicht: (2024)
Simulating Students or Sycophantic Problem Solving? On Misconception Faithfulness of LLM Simulators
von: Do, Heejin, et al.
Veröffentlicht: (2026)
von: Do, Heejin, et al.
Veröffentlicht: (2026)
Towards Aligning Language Models with Textual Feedback
von: Lloret, Saüc Abadal, et al.
Veröffentlicht: (2024)
von: Lloret, Saüc Abadal, et al.
Veröffentlicht: (2024)
Factual Self-Awareness in Language Models: Representation, Robustness, and Scaling
von: Tamoyan, Hovhannes, et al.
Veröffentlicht: (2025)
von: Tamoyan, Hovhannes, et al.
Veröffentlicht: (2025)
Probing for Arithmetic Errors in Language Models
von: Sun, Yucheng, et al.
Veröffentlicht: (2025)
von: Sun, Yucheng, et al.
Veröffentlicht: (2025)
Can Large Language Models Infer Causation from Correlation?
von: Jin, Zhijing, et al.
Veröffentlicht: (2023)
von: Jin, Zhijing, et al.
Veröffentlicht: (2023)
Stepwise Self-Consistent Mathematical Reasoning with Large Language Models
von: Zhao, Zilong, et al.
Veröffentlicht: (2024)
von: Zhao, Zilong, et al.
Veröffentlicht: (2024)
Variational Learning is Effective for Large Deep Networks
von: Shen, Yuesong, et al.
Veröffentlicht: (2024)
von: Shen, Yuesong, et al.
Veröffentlicht: (2024)
Towards Automated Error Discovery: A Study in Conversational AI
von: Petrak, Dominic, et al.
Veröffentlicht: (2025)
von: Petrak, Dominic, et al.
Veröffentlicht: (2025)
Learning to Reason Efficiently with A* Post-Training
von: Opedal, Andreas, et al.
Veröffentlicht: (2026)
von: Opedal, Andreas, et al.
Veröffentlicht: (2026)
Confidence Regulation Neurons in Language Models
von: Stolfo, Alessandro, et al.
Veröffentlicht: (2024)
von: Stolfo, Alessandro, et al.
Veröffentlicht: (2024)
CLadder: Assessing Causal Reasoning in Language Models
von: Jin, Zhijing, et al.
Veröffentlicht: (2023)
von: Jin, Zhijing, et al.
Veröffentlicht: (2023)
Token Weighting for Long-Range Language Modeling
von: Helm, Falko, et al.
Veröffentlicht: (2025)
von: Helm, Falko, et al.
Veröffentlicht: (2025)
Stepwise Perplexity-Guided Refinement for Efficient Chain-of-Thought Reasoning in Large Language Models
von: Cui, Yingqian, et al.
Veröffentlicht: (2025)
von: Cui, Yingqian, et al.
Veröffentlicht: (2025)
SPARE: Single-Pass Annotation with Reference-Guided Evaluation for Automatic Process Supervision and Reward Modelling
von: Rizvi, Md Imbesat Hassan, et al.
Veröffentlicht: (2025)
von: Rizvi, Md Imbesat Hassan, et al.
Veröffentlicht: (2025)
PATS: Personality-Aware Teaching Strategies with Large Language Model Tutors
von: Rooein, Donya, et al.
Veröffentlicht: (2026)
von: Rooein, Donya, et al.
Veröffentlicht: (2026)
Are Language Models Efficient Reasoners? A Perspective from Logic Programming
von: Opedal, Andreas, et al.
Veröffentlicht: (2025)
von: Opedal, Andreas, et al.
Veröffentlicht: (2025)
MathGAP: Out-of-Distribution Evaluation on Problems with Arbitrarily Complex Proofs
von: Opedal, Andreas, et al.
Veröffentlicht: (2024)
von: Opedal, Andreas, et al.
Veröffentlicht: (2024)
Do Language Models Exhibit the Same Cognitive Biases in Problem Solving as Human Learners?
von: Opedal, Andreas, et al.
Veröffentlicht: (2024)
von: Opedal, Andreas, et al.
Veröffentlicht: (2024)
Forward-Backward Reasoning in Large Language Models for Mathematical Verification
von: Jiang, Weisen, et al.
Veröffentlicht: (2023)
von: Jiang, Weisen, et al.
Veröffentlicht: (2023)
Auditing Language Model Unlearning via Information Decomposition
von: Goel, Anmol, et al.
Veröffentlicht: (2026)
von: Goel, Anmol, et al.
Veröffentlicht: (2026)
Pointwise Mutual Information as a Performance Gauge for Retrieval-Augmented Generation
von: Liu, Tianyu, et al.
Veröffentlicht: (2024)
von: Liu, Tianyu, et al.
Veröffentlicht: (2024)
FIRE: Fact-checking with Iterative Retrieval and Verification
von: Xie, Zhuohan, et al.
Veröffentlicht: (2024)
von: Xie, Zhuohan, et al.
Veröffentlicht: (2024)
Improving LoRA with Variational Learning
von: Cong, Bai, et al.
Veröffentlicht: (2025)
von: Cong, Bai, et al.
Veröffentlicht: (2025)
Variation in Verification: Understanding Verification Dynamics in Large Language Models
von: Zhou, Yefan, et al.
Veröffentlicht: (2025)
von: Zhou, Yefan, et al.
Veröffentlicht: (2025)
Autoformalizing Natural Language to First-Order Logic: A Case Study in Logical Fallacy Detection
von: Lalwani, Abhinav, et al.
Veröffentlicht: (2024)
von: Lalwani, Abhinav, et al.
Veröffentlicht: (2024)
Dense SAE Latents Are Features, Not Bugs
von: Sun, Xiaoqing, et al.
Veröffentlicht: (2025)
von: Sun, Xiaoqing, et al.
Veröffentlicht: (2025)
SMART: Self-learning Meta-strategy Agent for Reasoning Tasks
von: Liu, Rongxing, et al.
Veröffentlicht: (2024)
von: Liu, Rongxing, et al.
Veröffentlicht: (2024)
Stepwise Alignment for Constrained Language Model Policy Optimization
von: Wachi, Akifumi, et al.
Veröffentlicht: (2024)
von: Wachi, Akifumi, et al.
Veröffentlicht: (2024)
Evaluating Mathematical Reasoning of Large Language Models: A Focus on Error Identification and Correction
von: Li, Xiaoyuan, et al.
Veröffentlicht: (2024)
von: Li, Xiaoyuan, et al.
Veröffentlicht: (2024)
Socratic Reasoning Improves Positive Text Rewriting
von: Goel, Anmol, et al.
Veröffentlicht: (2024)
von: Goel, Anmol, et al.
Veröffentlicht: (2024)
Stepwise Penalization for Length-Efficient Chain-of-Thought Reasoning
von: Li, Xintong, et al.
Veröffentlicht: (2026)
von: Li, Xintong, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
MathTutorBench: A Benchmark for Measuring Open-ended Pedagogical Capabilities of LLM Tutors
von: Macina, Jakub, et al.
Veröffentlicht: (2025) -
From Problem-Solving to Teaching Problem-Solving: Aligning LLMs with Pedagogy using Reinforcement Learning
von: Dinucu-Jianu, David, et al.
Veröffentlicht: (2025) -
Towards the Pedagogical Steering of Large Language Models for Tutoring: A Case Study with Modeling Productive Failure
von: Puech, Romain, et al.
Veröffentlicht: (2024) -
Uncertainty-Aware Decoding with Minimum Bayes Risk
von: Daheim, Nico, et al.
Veröffentlicht: (2025) -
Model Merging by Uncertainty-Based Gradient Matching
von: Daheim, Nico, et al.
Veröffentlicht: (2023)