paper-with-me

홈 › Papers

Benchmarking Scientific Machine Learning Models for Air Quality Data

2026-03-22 · Khawja Imran Masud, Venkata Sai Rahul Unnam, Sahara Ali arxiv

Accurate air quality index (AQI) forecasting is essential for the protecting public health in rapidly growing urban regions, and the practical model evaluation and selection are often challenged by the lack of rigorous, region-specific benchmarking on standardized datasets. Physics-guided machine learning and deep learning models could be a good and effective solution to resolve such issues with more accurate and efficient AQI forecasting. This research study presents an explainable and comprehensive benchmark that enables a guideline and proposed physics-guided best model by benchmarking classical time-series, machine-learning, and deep-learning approaches for multi-horizon AQI forecasting in North Texas (Dallas County). Using publicly available U.S. Environmental Protection Agency (EPA) daily observations of air quality data from 2022 to 2024, we curate city-level time series for PM2.5 and O3 by aggregating station measurements and constructing lag-wise forecasting datasets for LAG in {1,7,14,30} days. For benchmarking the best model, linear regression (LR), SARIMAX, multilayer perceptrons (MLP), and LSTM networks are evaluated with the proposed physics-guided variants (MLP+Physics and LSTM+Physics) that incorporate the EPA breakpoint-based AQI formulation as a consistency constraint through a weighted loss. Experiments using chronological train-test splits and error metrics MAE, RMSE showed that deep-learning models outperform simpler baselines, while physics guidance improves stability and yields physically consistent pollutant with AQI relationships, with the largest benefits observed for short-horizon prediction and for PM2.5 and O3. Overall, the results provide a practical reference for selecting AQI forecasting models in North Texas and clarify when lightweight physics constraints meaningfully improve predictive performance across pollutants and forecast horizons.

📄 PDF Abstract BibTeX arXiv:2603.21039

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Scientific Machine Learning Benchmarks

2021-10-25 · Jeyan Thiyagalingam, Mallikarjun Shankar, Geoffrey Fox, Tony Hey

The breakthrough in Deep Learning neural networks has transformed the use of AI and machine learning technologies for the analysis of very large experimental datasets. These datasets are typically generated by large-scal…

BenchmarkingBIG-bench Machine Learning

An MLCommons Scientific Benchmarks Ontology

2025-11-06 · Ben Hawks, Gregor von Laszewski, Matthew D. Sinclair, Marco Colombo 외 arxiv

Scientific machine learning research spans diverse domains and data modalities, yet existing benchmark efforts remain siloed and lack standardization. This makes novel and transformative applications of machine learning …

Towards a Benchmark for Scientific Understanding in Humans and Machines

2023-04-20 · Kristian Gonzalez Barman, Sascha Caron, Tom Claassen, Henk de Regt

Scientific understanding is a fundamental goal of science, allowing us to explain the world. There is currently no good way to measure the scientific understanding of agents, whether these be humans or Artificial Intelli…

BenchmarkingInformation RetrievalPhilosophyRetrieval

Does AI for science need another ImageNet Or totally different benchmarks? A case study of machine learning force fields

2023-08-11 · Yatao Li, Wanling Gao, Lei Wang, Lixin Sun 외

AI for science (AI4S) is an emerging research field that aims to enhance the accuracy and speed of scientific computing tasks using machine learning methods. Traditional AI benchmarking methods struggle to adapt to the u…

Benchmarking

The Benchmarking Epistemology: Construct Validity for Evaluating Machine Learning Models

2025-10-27 · Timo Freiesleben, Sebastian Zezulka arxiv

Predictive benchmarking, the evaluation of machine learning models based on predictive performance and competitive ranking, is a central epistemic practice in machine learning research and an increasingly prominent metho…

Image Classification