paper-with-me

홈 › Papers

Comparative Evaluation of Explainable Machine Learning Versus Linear Regression for Predicting County-Level Lung Cancer Mortality Rate in the United States

2025-12-10 · Soheil Hashtarkhani, Brianna M. White, Benyamin Hoseini, David L. Schwartz, Arash Shaban-Nejad arxiv

Lung cancer (LC) is a leading cause of cancer-related mortality in the United States. Accurate prediction of LC mortality rates is crucial for guiding targeted interventions and addressing health disparities. Although traditional regression-based models have been commonly used, explainable machine learning models may offer enhanced predictive accuracy and deeper insights into the factors influencing LC mortality. This study applied three models: random forest (RF), gradient boosting regression (GBR), and linear regression (LR) to predict county-level LC mortality rates across the United States. Model performance was evaluated using R-squared and root mean squared error (RMSE). Shapley Additive Explanations (SHAP) values were used to determine variable importance and their directional impact. Geographic disparities in LC mortality were analyzed through Getis-Ord (Gi*) hotspot analysis. The RF model outperformed both GBR and LR, achieving an R2 value of 41.9% and an RMSE of 12.8. SHAP analysis identified smoking rate as the most important predictor, followed by median home value and the percentage of the Hispanic ethnic population. Spatial analysis revealed significant clusters of elevated LC mortality in the mid-eastern counties of the United States. The RF model demonstrated superior predictive performance for LC mortality rates, emphasizing the critical roles of smoking prevalence, housing values, and the percentage of Hispanic ethnic population. These findings offer valuable actionable insights for designing targeted interventions, promoting screening, and addressing health disparities in regions most affected by LC in the United States.

📄 PDF Abstract BibTeX arXiv:2512.17934

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Explaining Deep Neural Networks and Beyond: A Review of Methods and Applications

2020-03-17 · Wojciech Samek, Grégoire Montavon, Sebastian Lapuschkin, Christopher J. Anders 외

With the broader and highly successful usage of machine learning in industry and the sciences, there has been a growing demand for Explainable AI. Interpretability and explanation methods for gaining a better understandi…

BIG-bench Machine LearningInterpretable Machine Learning

Comparative analysis of machine learning methods for active flow control

2022-02-23 · Fabio Pino, Lorenzo Schena, Jean Rabault, Miguel A. Mendez

Machine learning frameworks such as Genetic Programming (GP) and Reinforcement Learning (RL) are gaining popularity in flow control. This work presents a comparative analysis of the two, bench-marking some of their most …

Bayesian OptimizationBIG-bench Machine Learningglobal-optimizationReinforcement Learning (RL)

GWRBoost:A geographically weighted gradient boosting method for explainable quantification of spatially-varying relationships

2022-12-12 · Han Wang, Zhou Huang, Ganmin Yin, Yi Bao 외

The geographically weighted regression (GWR) is an essential tool for estimating the spatial variation of relationships between dependent and independent variables in geographical contexts. However, GWR suffers from the …

parameter estimationregression

Explainable AI for Comparative Analysis of Intrusion Detection Models

2024-06-14 · Pap M. Corea, Yongxin Liu, Jian Wang, Shuteng Niu 외

Explainable Artificial Intelligence (XAI) has become a widely discussed topic, the related technologies facilitate better understanding of conventional black-box models like Random Forest, Neural Networks and etc. Howeve…

Explainable artificial intelligenceExplainable Artificial Intelligence (XAI)Feature EngineeringIntrusion Detection+2

"This Suits You the Best": Query Focused Comparative Explainable Summarization

2025-07-07 · Arnav Attri, Anuj Attri, Pushpak Bhattacharyya, Suman Banerjee 외 arxiv

Product recommendations inherently involve comparisons, yet traditional opinion summarization often fails to provide holistic comparative insights. We propose the novel task of generating Query-Focused Comparative Explai…