Ask-Before-Detection: Identifying and Mitigating Conformity Bias in LLM-Powered Error Detector for Math Word Problem Solutions
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Hang, Xu, Tianlong, Yang, Kaiqi, Chu, Yucheng, Chen, Yanling, Song, Yichi, Wen, Qingsong, Liu, Hui |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Iterative LLM-Based Generation and Refinement of Distracting Conditions in Math Word Problems
di: Yang, Kaiqi, et al.
Pubblicazione: (2025)
di: Yang, Kaiqi, et al.
Pubblicazione: (2025)
Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
di: Li, Hang, et al.
Pubblicazione: (2025)
di: Li, Hang, et al.
Pubblicazione: (2025)
Automate Knowledge Concept Tagging on Math Questions with LLMs
di: Li, Hang, et al.
Pubblicazione: (2024)
di: Li, Hang, et al.
Pubblicazione: (2024)
Knowledge Tagging System on Math Questions via LLMs with Flexible Demonstration Retriever
di: Li, Hang, et al.
Pubblicazione: (2024)
di: Li, Hang, et al.
Pubblicazione: (2024)
Can MLLMs Read Students' Minds? Unpacking Multimodal Error Analysis in Handwritten Math
di: Song, Dingjie, et al.
Pubblicazione: (2026)
di: Song, Dingjie, et al.
Pubblicazione: (2026)
A LLM-Powered Automatic Grading Framework with Human-Level Guidelines Optimization
di: Chu, Yucheng, et al.
Pubblicazione: (2024)
di: Chu, Yucheng, et al.
Pubblicazione: (2024)
Multimodal AI Teacher: Integrating Edge Computing and Reasoning Models for Enhanced Student Error Analysis
di: Tianlong Xu, et al.
Pubblicazione: (2025)
di: Tianlong Xu, et al.
Pubblicazione: (2025)
From Correctness to Comprehension: AI Agents for Personalized Error Diagnosis in Education
di: Zhang, Yi-Fan, et al.
Pubblicazione: (2025)
di: Zhang, Yi-Fan, et al.
Pubblicazione: (2025)
Beyond Partisan Leaning: A Comparative Analysis of Political Bias in Large Language Models
di: Peng, Tai-Quan, et al.
Pubblicazione: (2024)
di: Peng, Tai-Quan, et al.
Pubblicazione: (2024)
LLM-based Automated Grading with Human-in-the-Loop
di: Chu, Yucheng, et al.
Pubblicazione: (2025)
di: Chu, Yucheng, et al.
Pubblicazione: (2025)
Error Classification of Large Language Models on Math Word Problems: A Dynamically Adaptive Framework
di: Sun, Yuhong, et al.
Pubblicazione: (2025)
di: Sun, Yuhong, et al.
Pubblicazione: (2025)
AI-Driven Virtual Teacher for Enhanced Educational Efficiency: Leveraging Large Pretrain Models for Autonomous Error Analysis and Correction
di: Xu, Tianlong, et al.
Pubblicazione: (2024)
di: Xu, Tianlong, et al.
Pubblicazione: (2024)
A LLM-Driven Multi-Agent Systems for Professional Development of Mathematics Teachers
di: Yang, Kaiqi, et al.
Pubblicazione: (2025)
di: Yang, Kaiqi, et al.
Pubblicazione: (2025)
Before Making the Ask …
Pubblicazione: (2024)
Pubblicazione: (2024)
Knowledge Tagging with Large Language Model based Multi-Agent System
di: Li, Hang, et al.
Pubblicazione: (2024)
di: Li, Hang, et al.
Pubblicazione: (2024)
MathAgent: Leveraging a Mixture-of-Math-Agent Framework for Real-World Multimodal Mathematical Error Detection
di: Yan, Yibo, et al.
Pubblicazione: (2025)
di: Yan, Yibo, et al.
Pubblicazione: (2025)
Cutting Through the Noise: Boosting LLM Performance on Math Word Problems
di: Anantheswaran, Ujjwala, et al.
Pubblicazione: (2024)
di: Anantheswaran, Ujjwala, et al.
Pubblicazione: (2024)
Adversarial Math Word Problem Generation
di: Xie, Roy, et al.
Pubblicazione: (2024)
di: Xie, Roy, et al.
Pubblicazione: (2024)
Ask for No Gift Before It's Time
Pubblicazione: (2024)
Pubblicazione: (2024)
Template-Driven LLM-Paraphrased Framework for Tabular Math Word Problem Generation
di: Kang, Xiaoqiang, et al.
Pubblicazione: (2024)
di: Kang, Xiaoqiang, et al.
Pubblicazione: (2024)
Optimizing In-Context Demonstrations for LLM-based Automated Grading
di: Chu, Yucheng, et al.
Pubblicazione: (2026)
di: Chu, Yucheng, et al.
Pubblicazione: (2026)
MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems?
di: Zhang, Renrui, et al.
Pubblicazione: (2024)
di: Zhang, Renrui, et al.
Pubblicazione: (2024)
Confusion-Aware Rubric Optimization for LLM-based Automated Grading
di: Chu, Yucheng, et al.
Pubblicazione: (2026)
di: Chu, Yucheng, et al.
Pubblicazione: (2026)
Fill in the Blank: Exploring and Enhancing LLM Capabilities for Backward Reasoning in Math Word Problems
di: Deb, Aniruddha, et al.
Pubblicazione: (2023)
di: Deb, Aniruddha, et al.
Pubblicazione: (2023)
Benchmarking Hallucination in Large Language Models based on Unanswerable Math Word Problem
di: Sun, Yuhong, et al.
Pubblicazione: (2024)
di: Sun, Yuhong, et al.
Pubblicazione: (2024)
What Makes Math Word Problems Challenging for LLMs?
di: Srivatsa, KV Aditya, et al.
Pubblicazione: (2024)
di: Srivatsa, KV Aditya, et al.
Pubblicazione: (2024)
Expression Syntax Information Bottleneck for Math Word Problems
di: Xiong, Jing, et al.
Pubblicazione: (2023)
di: Xiong, Jing, et al.
Pubblicazione: (2023)
Self-consistent Reasoning For Solving Math Word Problems
di: Xiong, Jing, et al.
Pubblicazione: (2022)
di: Xiong, Jing, et al.
Pubblicazione: (2022)
Does LLM Focus on the Right Words? Mitigating Context Bias in LLM-based Recommenders
di: Wang, Bohao, et al.
Pubblicazione: (2025)
di: Wang, Bohao, et al.
Pubblicazione: (2025)
TreeCut: A Synthetic Unanswerable Math Word Problem Dataset for LLM Hallucination Evaluation
di: Ouyang, Jialin
Pubblicazione: (2025)
di: Ouyang, Jialin
Pubblicazione: (2025)
Identifying and Upweighting Power-Niche Users to Mitigate Popularity Bias in Recommendations
di: Liu, David, et al.
Pubblicazione: (2025)
di: Liu, David, et al.
Pubblicazione: (2025)
We Need Knowledge Distillation for Solving Math Word Problems
di: Shen, Zhenquan, et al.
Pubblicazione: (2025)
di: Shen, Zhenquan, et al.
Pubblicazione: (2025)
Structured Reasoning with Tree-of-Thoughts for Bengali Math Word Problems
di: Mahmood, Aurprita, et al.
Pubblicazione: (2025)
di: Mahmood, Aurprita, et al.
Pubblicazione: (2025)
EDUMATH: Generating Standards-aligned Educational Math Word Problems
di: Christ, Bryan R., et al.
Pubblicazione: (2025)
di: Christ, Bryan R., et al.
Pubblicazione: (2025)
Can LLMs Solve longer Math Word Problems Better?
di: Xu, Xin, et al.
Pubblicazione: (2024)
di: Xu, Xin, et al.
Pubblicazione: (2024)
Augmenting Math Word Problems via Iterative Question Composing
di: Liu, Haoxiong, et al.
Pubblicazione: (2024)
di: Liu, Haoxiong, et al.
Pubblicazione: (2024)
Wording a Million‐Dollar Ask
Pubblicazione: (2024)
Pubblicazione: (2024)
Identifying and Mitigating Systemic Measurement Bias in Production LLM Inference Benchmarks
di: Chandrasekar, Ashok, et al.
Pubblicazione: (2026)
di: Chandrasekar, Ashok, et al.
Pubblicazione: (2026)
Mitigating Gender Bias in Contextual Word Embeddings
di: Yarrabelly, Navya, et al.
Pubblicazione: (2024)
di: Yarrabelly, Navya, et al.
Pubblicazione: (2024)
LLM The Genius Paradox: A Linguistic and Math Expert's Struggle with Simple Word-based Counting Problems
di: Xu, Nan, et al.
Pubblicazione: (2024)
di: Xu, Nan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Iterative LLM-Based Generation and Refinement of Distracting Conditions in Math Word Problems
di: Yang, Kaiqi, et al.
Pubblicazione: (2025) -
Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
di: Li, Hang, et al.
Pubblicazione: (2025) -
Automate Knowledge Concept Tagging on Math Questions with LLMs
di: Li, Hang, et al.
Pubblicazione: (2024) -
Knowledge Tagging System on Math Questions via LLMs with Flexible Demonstration Retriever
di: Li, Hang, et al.
Pubblicazione: (2024) -
Can MLLMs Read Students' Minds? Unpacking Multimodal Error Analysis in Handwritten Math
di: Song, Dingjie, et al.
Pubblicazione: (2026)