ControlMath: Controllable Data Generation Promotes Math Generalist Models
Utilizing large language models (LLMs) for data augmentation has yielded encouraging results in mathematical reasoning. However, these approaches face constraints in problem diversity, potentially restricting them to in-domain/distribution data generation. To this end, we propose ControlMath, an iterative method involving an equation-generator module and two LLM-based agents. The module creates diverse equations, which the Problem-Crafter agent then transforms into math word problems. The Reverse-Agent filters and selects high-quality data, adhering to the "less is more" principle, achieving better results with fewer data points. This approach enables the generation of diverse math problems, not limited to specific domains or distributions. As a result, we collect ControlMathQA, which involves 190k math word problems. Extensive results prove that combining our dataset with in-domain datasets like GSM8K can help improve the model's mathematical ability to generalize, leading to improved performances both within and beyond specific domains.
Code (0)
등록된 구현이 없습니다.
Tasks
Data AugmentationDiversityGSM8KMathMathematical ReasoningSimilar Papers 제목 키워드 기반
HumanDiffusion: a Coarse-to-Fine Alignment Diffusion Framework for Controllable Text-Driven Person Image Generation
Text-driven person image generation is an emerging and challenging task in cross-modality image generation. Controllable person image generation promotes a wide range of applications such as digital human interaction and…
Image GenerationRetrievalSentenceVirtual Try-on$\mathtt{M^3VIR}$: A Large-Scale Multi-Modality Multi-View Synthesized Benchmark Dataset for Image Restoration and Content Creation
The gaming and entertainment industry is rapidly evolving, driven by immersive experiences and the integration of generative AI (GAI) technologies. Training such models effectively requires large-scale datasets that capt…
Novel View SynthesisImage RestorationVideo GenerationPlay to Generalize: Learning to Reason Through Game Play
Developing generalizable reasoning capabilities in multimodal large language models (MLLMs) remains challenging. Motivated by cognitive science literature suggesting that gameplay promotes transferable cognitive skills, …
Domain GeneralizationMathMultimodal ReasoningReinforcement Learning (RL)DualGenerator: Information Interaction-based Generative Network for Point Cloud Completion
Point cloud completion estimates complete shapes from incomplete point clouds to obtain higher-quality point cloud data. Most existing methods only consider global object features, ignoring spatial and semantic informati…
Point Cloud CompletionA mathematical modelling portrait of Wnt signalling in early vertebrate embryogenesis
There are two phases of Wnt signalling in early vertebrate embryogenesis: very early, maternal Wnt signalling promotes dorsal development, and slightly later, zygotic Wnt signalling promotes ventral and lateral mesoderm …