paper-with-me

Papers

MatFormBench: A Benchmarking Evaluation Framework for Target-Driven Materials Formulation

2026-05-26 · Linhan Wu, Chenxi Wang, Chuhan Yang, Zhengwei Yang, Yuyang Liu arxiv

Inverse design of materials has significantly advanced target-driven formulation optimization, yet existing materials machine learning benchmarks remain limited to forward property prediction, failing to systematically evaluate inverse optimization and generation algorithms, a critical gap that hinders the progress of target-driven materials design. To address this limitation, we propose MatFormBench, a novel benchmarking ecosystem tailored to evaluate and guide generative strategies for target-driven formulation. MatFormBench integrates a physics-driven formulation generation scheme to generate synthetic samples that faithfully emulate realistic materials structure-property response relationships, complemented by five escalating difficulty levels to quantify the complexity of these relationships. To rigorously assess algorithm performance, we further propose MatFormScore, a multi-dimensional metric that comprehensively quantifies performance across five critical axes: target success, search efficiency, exploratory capacity, robustness, and stability. We validate MatFormBench by evaluating 39 diverse inverse design algorithms, covering classical surrogate-assisted black-box search, state-of-the-art deep generative models, and increasingly popular Large Language Model (LLM)-based recommendation strategies. Across 1170 standardized algorithm-task evaluations, diffusion-based models demonstrate the strongest overall performance, while Variational Autoencoder (VAE)-based and Genetic Algorithm (GA)-based methods exhibit distinct advantages in specific scenarios. By establishing a unified evaluation standard for target-driven materials formulation, MatFormBench enables reproducible benchmarking, principled algorithm comparison, and diagnostic analysis of inverse design strategies, providing a foundational tool for advancing materials inverse design.

📄 PDF Abstract BibTeX arXiv:2605.26741

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

WelQrate: Defining the Gold Standard in Small Molecule Drug Discovery Benchmarking

2024-11-14 · Yunchao, Liu, Ha Dong, Xin Wang 외

While deep learning has revolutionized computer-aided drug discovery, the AI community has predominantly focused on model innovation and placed less emphasis on establishing best benchmarking practices. We posit that wit…

BenchmarkingDrug Discovery

Benchmarking Synthetic Tabular Data: A Multi-Dimensional Evaluation Framework

2025-04-02 · Andrey Sidorenko, Michael Platzer, Mario Scriminaci, Paul Tiwald

Evaluating the quality of synthetic data remains a key challenge for ensuring privacy and utility in data-driven research. In this work, we present an evaluation framework that quantifies how well synthetic data replicat…

BenchmarkingSynthetic Data Generation

Towards a more inductive world for drug repurposing approaches

2023-11-21 · Jesus de la Fuente, Guillermo Serrano, Uxía Veleiro, Mikel Casals 외

Drug-target interaction (DTI) prediction is a challenging, albeit essential task in drug repurposing. Learning on graph models have drawn special attention as they can significantly reduce drug repurposing costs and time…

BenchmarkingPrediction

AutoJudger: An Agent-Driven Framework for Efficient Benchmarking of MLLMs

2025-05-27 · Xuanwen Ding, Chengjun Pan, Zejun Li, Jiwen Zhang 외

Evaluating multimodal large language models (MLLMs) is increasingly expensive, as the growing size and cross-modality complexity of benchmarks demand significant scoring efforts. To tackle with this difficulty, we introd…

BenchmarkingQuestion Selection

Accelerating Reproducible Research in Synthetic EHR Generation

2026-06-05 · Jalen Jiang, Chufan Gao, Ethan Rasmussen, Stephen Z. Xie 외 arxiv

The generation of high-fidelity synthetic Electronic Health Records (EHR) is crucial for advancing medical research while preserving patient privacy. However, head-to-head comparison of existing generative models is hind…