paper-with-me

홈 › Papers

Scientific Image Synthesis: Benchmarking, Methodologies, and Downstream Utility

2026-01-17 · Honglin Lin, Chonghan Qin, Zheng Liu, Qizhi Pei, Yu Li, Zhanping Zhong, Xin Gao, Yanfeng Wang, Conghui He, Lijun Wu arxiv

While synthetic data has proven effective for improving scientific reasoning in the text domain, multimodal reasoning remains constrained by the difficulty of synthesizing scientifically rigorous images. Existing Text-to-Image (T2I) models often produce outputs that are visually plausible yet scientifically incorrect, resulting in a persistent visual-logic divergence that limits their value for downstream reasoning. Motivated by recent advances in next-generation T2I models, we conduct a systematic study of scientific image synthesis across generation paradigms, evaluation, and downstream use. We analyze both direct pixel-based generation and programmatic synthesis, and propose ImgCoder, a logic-driven framework that follows an explicit "understand - plan - code" workflow to improve structural precision. To rigorously assess scientific correctness, we introduce SciGenBench, which evaluates generated images based on information utility and logical validity. Our evaluation reveals systematic failure modes in pixel-based models and highlights a fundamental expressiveness-precision trade-off. Finally, we show that fine-tuning Large Multimodal Models (LMMs) on rigorously verified synthetic scientific images yields consistent reasoning gains, with potential scaling trends analogous to the text domain, validating high-fidelity scientific synthesis as a viable path to unlocking massive multimodal reasoning capabilities.

📄 PDF Abstract BibTeX arXiv:2601.17027

Code (0)

등록된 구현이 없습니다.

Tasks

Multimodal Reasoning

Similar Papers 제목 키워드 기반

Does AI for science need another ImageNet Or totally different benchmarks? A case study of machine learning force fields

2023-08-11 · Yatao Li, Wanling Gao, Lei Wang, Lixin Sun 외

AI for science (AI4S) is an emerging research field that aims to enhance the accuracy and speed of scientific computing tasks using machine learning methods. Traditional AI benchmarking methods struggle to adapt to the u…

Benchmarking

The evolution of AI from image interpretation toward scientific inference in nanoparticle electron microscopy

2026-07-11 · Evropi Toulkeridou, Jiafei Li, Leonardo Lari, Panagiotis Grammatikopoulos arxiv

Artificial intelligence (AI) is transforming electron microscopy by enabling quantitative analysis of increasingly large and complex datasets for nanoparticle characterization. Recent advances in machine learning (ML) an…

Self-Supervised Learning

Q-space Conditioned Translation Networks for Directional Synthesis of Diffusion Weighted Images from Multi-modal Structural MRI

2021-06-24 · Mengwei Ren, Heejong Kim, Neel Dey, Guido Gerig

Current deep learning approaches for diffusion MRI modeling circumvent the need for densely-sampled diffusion-weighted images (DWIs) by directly predicting microstructural indices from sparsely-sampled DWIs. However, the…

Diffusion MRITranslation

Generative AI for Synthetic Data Across Multiple Medical Modalities: A Systematic Review of Recent Developments and Challenges

2024-06-27 · Mahmoud Ibrahim, Yasmina Al Khalil, Sina Amirrajab, Chang Sun 외

This paper presents a comprehensive systematic review of generative models (GANs, VAEs, DMs, and LLMs) used to synthesize various medical data types, including imaging (dermoscopic, mammographic, ultrasound, CT, MRI, and…

BenchmarkingClinical Knowledge

Analyzing the Feature Extractor Networks for Face Image Synthesis

2024-06-04 · Erdi Sarıtaş, Hazim Kemal Ekenel

Advancements like Generative Adversarial Networks have attracted the attention of researchers toward face image synthesis to generate ever more realistic images. Thereby, the need for the evaluation criteria to assess th…

BenchmarkingImage Generation