paper-with-me

Papers

AQFusionNet: Multimodal Deep Learning for Air Quality Index Prediction with Imagery and Sensor Data

2025-08-30 · Koushik Ahmed Kushal, Abdullah Al Mamun arxiv

Air pollution monitoring in resource-constrained regions remains challenging due to sparse sensor deployment and limited infrastructure. This work introduces AQFusionNet, a multimodal deep learning framework for robust Air Quality Index (AQI) prediction. The framework integrates ground-level atmospheric imagery with pollutant concentration data using lightweight CNN backbones (MobileNetV2, ResNet18, EfficientNet-B0). Visual and sensor features are combined through semantically aligned embedding spaces, enabling accurate and efficient prediction. Experiments on more than 8,000 samples from India and Nepal demonstrate that AQFusionNet consistently outperforms unimodal baselines, achieving up to 92.02% classification accuracy and an RMSE of 7.70 with the EfficientNet-B0 backbone. The model delivers an 18.5% improvement over single-modality approaches while maintaining low computational overhead, making it suitable for deployment on edge devices. AQFusionNet provides a scalable and practical solution for AQI monitoring in infrastructure-limited environments, offering robust predictive capability even under partial sensor availability.

📄 PDF Abstract BibTeX arXiv:2509.00353

Code (0)

등록된 구현이 없습니다.

Tasks

Multimodal Deep Learning

Similar Papers 제목 키워드 기반

Predicting air quality via multimodal AI and satellite imagery

2022-11-01 · Andrew Rowley, Oktay Karakuş

Climate change may be classified as the most important environmental problem that the Earth is currently facing, and affects all living species on Earth. Given that air-quality monitoring stations are typically ground-ba…

Spatiotemporal Air Quality Mapping in Urban Areas Using Sparse Sensor Data, Satellite Imagery, Meteorological Factors, and Spatial Features

2025-01-20 · Osama Ahmad, Zubair Khalid, Muhammad Tahir, Momin Uppal

Monitoring air pollution is crucial for protecting human health from exposure to harmful substances. Traditional methods of air quality monitoring, such as ground-based sensors and satellite-based remote sensing, face li…

StreetviewLLM: Extracting Geographic Information Using a Chain-of-Thought Multimodal Large Language Model

2024-11-19 · Zongrong Li, Junhao Xu, Siqin Wang, Yifan Wu 외

Geospatial predictions are crucial for diverse fields such as disaster management, urban planning, and public health. Traditional machine learning methods often face limitations when handling unstructured or multi-modal …

Decision MakingLanguage ModelingLanguage ModellingLarge Language Model+3

Designing streetscapes from street-view imagery using diffusion models

2026-05-17 · Yuzhou Chen, Yuebing Liang, Lingqian Hu, Kailai Sun 외 arxiv

Street-view imagery (SVI) is widely used to quantify key indicators of urban environment, such as green- ery, sky, or road view indices. However, existing studies largely focus on measuring current streetscapes and rarel…

Scene Generation

WalkCLIP: Multimodal Learning for Urban Walkability Prediction

2025-11-26 · Shilong Xiang, JangHyeon Lee, Min Namgung, Yao-Yi Chiang arxiv

Urban walkability is a cornerstone of public health, sustainability, and quality of life. Traditional walkability assessments rely on surveys and field audits, which are costly and difficult to scale. Recent studies have…