Generating Pedagogically Meaningful Visuals for Math Word Problems: A New Benchmark and Analysis of Text-to-Image Models
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Junling, Rutkiewicz, Anna, Wang, April Yi, Sachan, Mrinmaya |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bridging Instead of Replacing Online Coding Communities with AI through Community-Enriched Chatbot Designs
by: Wang, Junling, et al.
Published: (2026)
by: Wang, Junling, et al.
Published: (2026)
AutoTutor meets Large Language Models: A Language Model Tutor with Rich Pedagogy and Guardrails
by: Chowdhury, Sankalan Pal, et al.
Published: (2024)
by: Chowdhury, Sankalan Pal, et al.
Published: (2024)
When Should Teachers Control AI Generation for Mathematics Visuals?
by: Li, Zhengxu, et al.
Published: (2026)
by: Li, Zhengxu, et al.
Published: (2026)
Benchmarking and Enhancing Text-to-Image Models for Generating Visual Representations in Early Arithmetic Education
by: Wang, Junling, et al.
Published: (2026)
by: Wang, Junling, et al.
Published: (2026)
RELIC: Investigating Large Language Model Responses using Self-Consistency
by: Cheng, Furui, et al.
Published: (2023)
by: Cheng, Furui, et al.
Published: (2023)
Can LLMs Model Incorrect Student Reasoning? A Case Study on Distractor Generation
by: Zengaffinen, Yanick, et al.
Published: (2026)
by: Zengaffinen, Yanick, et al.
Published: (2026)
LLM-based Cognitive Models of Students with Misconceptions
by: Sonkar, Shashank, et al.
Published: (2024)
by: Sonkar, Shashank, et al.
Published: (2024)
Word Synchronization Challenge: A Benchmark for Word Association Responses for Large Language Models
by: Cazalets, Tanguy, et al.
Published: (2025)
by: Cazalets, Tanguy, et al.
Published: (2025)
Product vs. Process: Exploring EFL Students' Editing of AI-Generated Text for Expository Writing
by: Woo, David James, et al.
Published: (2025)
by: Woo, David James, et al.
Published: (2025)
Can Vision-Language Models Solve Visual Math Equations?
by: Choudhury, Monjoy Narayan, et al.
Published: (2025)
by: Choudhury, Monjoy Narayan, et al.
Published: (2025)
DBox: Scaffolding Algorithmic Programming Learning through Learner-LLM Co-Decomposition
by: Ma, Shuai, et al.
Published: (2025)
by: Ma, Shuai, et al.
Published: (2025)
Emotionally Aware Moderation: The Potential of Emotion Monitoring in Shaping Healthier Social Media Conversations
by: Su, Xiaotian, et al.
Published: (2025)
by: Su, Xiaotian, et al.
Published: (2025)
ELI-Why: Evaluating the Pedagogical Utility of Language Model Explanations
by: Joshi, Brihi, et al.
Published: (2025)
by: Joshi, Brihi, et al.
Published: (2025)
UI Remix: Supporting UI Design Through Interactive Example Retrieval and Remixing
by: Wang, Junling, et al.
Published: (2026)
by: Wang, Junling, et al.
Published: (2026)
Many Ways to Be Fake: Benchmarking Fake News Detection Under Strategy-Driven AI Generation
by: Wang, Xinyu, et al.
Published: (2026)
by: Wang, Xinyu, et al.
Published: (2026)
VisEval: A Benchmark for Data Visualization in the Era of Large Language Models
by: Chen, Nan, et al.
Published: (2024)
by: Chen, Nan, et al.
Published: (2024)
Augmenting a Large Language Model with a Combination of Text and Visual Data for Conversational Visualization of Global Geospatial Data
by: Mena, Omar, et al.
Published: (2025)
by: Mena, Omar, et al.
Published: (2025)
Implicit Personalization in Language Models: A Systematic Study
by: Jin, Zhijing, et al.
Published: (2024)
by: Jin, Zhijing, et al.
Published: (2024)
Do Text-to-Vis Benchmarks Test Real Use of Visualisations?
by: Nguyen, Hy, et al.
Published: (2024)
by: Nguyen, Hy, et al.
Published: (2024)
Unknown Word Detection for English as a Second Language (ESL) Learners Using Gaze and Pre-trained Language Models
by: Ding, Jiexin, et al.
Published: (2025)
by: Ding, Jiexin, et al.
Published: (2025)
MathBuddy: A Multimodal System for Affective Math Tutoring
by: Kar, Debanjana, et al.
Published: (2025)
by: Kar, Debanjana, et al.
Published: (2025)
ChartInsighter: An Approach for Mitigating Hallucination in Time-series Chart Summary Generation with A Benchmark Dataset
by: Wang, Fen, et al.
Published: (2025)
by: Wang, Fen, et al.
Published: (2025)
MathVC: An LLM-Simulated Multi-Character Virtual Classroom for Mathematics Education
by: Yue, Murong, et al.
Published: (2024)
by: Yue, Murong, et al.
Published: (2024)
WordCraft: Scaffolding the Keyword Method for L2 Vocabulary Learning with Multimodal LLMs
by: Shao, Yuheng, et al.
Published: (2026)
by: Shao, Yuheng, et al.
Published: (2026)
Repairs in a Block World: A New Benchmark for Handling User Corrections with Multi-Modal Language Models
by: Chiyah-Garcia, Javier, et al.
Published: (2024)
by: Chiyah-Garcia, Javier, et al.
Published: (2024)
Towards the Pedagogical Steering of Large Language Models for Tutoring: A Case Study with Modeling Productive Failure
by: Puech, Romain, et al.
Published: (2024)
by: Puech, Romain, et al.
Published: (2024)
CPG-EVAL: A Multi-Tiered Benchmark for Evaluating the Chinese Pedagogical Grammar Competence of Large Language Models
by: Wang, Dong
Published: (2025)
by: Wang, Dong
Published: (2025)
Media of Langue: The Interface for Exploring Word Translation Network/Space
by: Muramoto, Goki, et al.
Published: (2023)
by: Muramoto, Goki, et al.
Published: (2023)
ReSpark: Leveraging Previous Data Reports as References to Generate New Reports with LLMs
by: Tian, Yuan, et al.
Published: (2025)
by: Tian, Yuan, et al.
Published: (2025)
Granuscore: A Reference-Free Measure of Granularity for Text Analysis and Question Answering
by: Ellinger, Lukas, et al.
Published: (2026)
by: Ellinger, Lukas, et al.
Published: (2026)
QE4PE: Word-level Quality Estimation for Human Post-Editing
by: Sarti, Gabriele, et al.
Published: (2025)
by: Sarti, Gabriele, et al.
Published: (2025)
An Actor-Critic Approach to Boosting Text-to-SQL Large Language Model
by: Zheng, Ziyang, et al.
Published: (2024)
by: Zheng, Ziyang, et al.
Published: (2024)
World Models for Math Story Problems
by: Opedal, Andreas, et al.
Published: (2023)
by: Opedal, Andreas, et al.
Published: (2023)
An Iterative Associative Memory Model for Empathetic Response Generation
by: Yang, Zhou, et al.
Published: (2024)
by: Yang, Zhou, et al.
Published: (2024)
JailbreakLens: Visual Analysis of Jailbreak Attacks Against Large Language Models
by: Feng, Yingchaojie, et al.
Published: (2024)
by: Feng, Yingchaojie, et al.
Published: (2024)
Word Clouds as Common Voices: LLM-Assisted Visualization of Participant-Weighted Themes in Qualitative Interviews
by: Colonel, Joseph T., et al.
Published: (2025)
by: Colonel, Joseph T., et al.
Published: (2025)
From Text to Trust: Empowering AI-assisted Decision Making with Adaptive LLM-powered Analysis
by: Li, Zhuoyan, et al.
Published: (2025)
by: Li, Zhuoyan, et al.
Published: (2025)
Interactive Discovery and Exploration of Visual Bias in Generative Text-to-Image Models
by: Eschner, Johannes, et al.
Published: (2025)
by: Eschner, Johannes, et al.
Published: (2025)
Can we Debias Social Stereotypes in AI-Generated Images? Examining Text-to-Image Outputs and User Perceptions
by: Barve, Saharsh, et al.
Published: (2025)
by: Barve, Saharsh, et al.
Published: (2025)
The Expressions of Depression and Anxiety in Chinese Psycho-counseling: Usage of First-person Singular Pronoun and Negative Emotional Words
by: Ma, Lizhi, et al.
Published: (2025)
by: Ma, Lizhi, et al.
Published: (2025)
Similar Items
-
Bridging Instead of Replacing Online Coding Communities with AI through Community-Enriched Chatbot Designs
by: Wang, Junling, et al.
Published: (2026) -
AutoTutor meets Large Language Models: A Language Model Tutor with Rich Pedagogy and Guardrails
by: Chowdhury, Sankalan Pal, et al.
Published: (2024) -
When Should Teachers Control AI Generation for Mathematics Visuals?
by: Li, Zhengxu, et al.
Published: (2026) -
Benchmarking and Enhancing Text-to-Image Models for Generating Visual Representations in Early Arithmetic Education
by: Wang, Junling, et al.
Published: (2026) -
RELIC: Investigating Large Language Model Responses using Self-Consistency
by: Cheng, Furui, et al.
Published: (2023)