paper-with-me

Papers

Generating Pedagogically Meaningful Visuals for Math Word Problems: A New Benchmark and Analysis of Text-to-Image Models

2025-06-04 · Junling Wang, Anna Rutkiewicz, April Yi Wang, Mrinmaya Sachan

Visuals are valuable tools for teaching math word problems (MWPs), helping young learners interpret textual descriptions into mathematical expressions before solving them. However, creating such visuals is labor-intensive and there is a lack of automated methods to support this process. In this paper, we present Math2Visual, an automatic framework for generating pedagogically meaningful visuals from MWP text descriptions. Math2Visual leverages a pre-defined visual language and a design space grounded in interviews with math teachers, to illustrate the core mathematical relationships in MWPs. Using Math2Visual, we construct an annotated dataset of 1,903 visuals and evaluate Text-to-Image (TTI) models for their ability to generate visuals that align with our design. We further fine-tune several TTI models with our dataset, demonstrating improvements in educational visual generation. Our work establishes a new benchmark for automated generation of pedagogically meaningful visuals and offers insights into key challenges in producing multimodal educational content, such as the misrepresentation of mathematical relationships and the omission of essential visual elements.

📄 PDF Abstract BibTeX arXiv:2506.03735

Code (1)

eth-lre/math2visual 공식 구현

Tasks

Math

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Benchmarking and Enhancing Text-to-Image Models for Generating Visual Representations in Early Arithmetic Education

2026-05-29 · Junling Wang, Boqi Chen, Heejin Do, Mubashara Akhtar 외 arxiv

AI systems are increasingly used to support educational content creation, yet it remains unclear whether they can generate outputs that faithfully represent the pedagogical concepts they are intended to teach. Thus, we i…

Image Generation

MAGMA-Edu: Multi-Agent Generative Multimodal Framework for Text-Diagram Educational Question Generation

2025-11-24 · Zhenyu Wu, Jian Li, Hua Huang arxiv

Educational illustrations play a central role in communicating abstract concepts, yet current multimodal large language models (MLLMs) remain limited in producing pedagogically coherent and semantically consistent educat…

Question GenerationImage Generation

I Do Not Understand What I Cannot Define: Automatic Question Generation With Pedagogically-Driven Content Selection

2021-10-08 · Tim Steuer, Anna Filighera, Tobias Meuser, Christoph Rensing

Most learners fail to develop deep text comprehension when reading textbooks passively. Posing questions about what learners have read is a well-established way of fostering their text comprehension. However, many textbo…

Learning TheoryQuestion GenerationQuestion-GenerationReading Comprehension

From Answers to Questions: EQGBench for Evaluating LLMs' Educational Question Generation

2025-08-05 · Chengliang Zhou, Mei Wang, Ting Zhang, Qiannan Zhu 외 arxiv

Large Language Models (LLMs) have demonstrated remarkable capabilities in mathematical problem-solving. However, the transition from providing answers to generating high-quality educational questions presents significant…

Question Generation

What Do I Hear? Generating Sounds for Visuals with ChatGPT

2023-11-09 · David Chuan-En Lin, Nikolas Martelaro

This short paper introduces a workflow for generating realistic soundscapes for visual media. In contrast to prior work, which primarily focus on matching sounds for on-screen visuals, our approach extends to suggesting …