Literature DB >> 33735646

Using a land use regression model with machine learning to estimate ground level PM2.5.

Pei-Yi Wong1, Hsiao-Yun Lee2, Yu-Cheng Chen3, Yu-Ting Zeng4, Yinq-Rong Chern4, Nai-Tzu Chen5, Shih-Chun Candice Lung6, Huey-Jen Su1, Chih-Da Wu7.   

Abstract

Ambient fine particulate matter (PM2.5) has been ranked as the sixth leading risk factor globally for death and disability. Modelling methods based on having access to a limited number of monitor stations are required for capturing PM2.5 spatial and temporal continuous variations with a sufficient resolution. This study utilized a land use regression (LUR) model with machine learning to assess the spatial-temporal variability of PM2.5. Daily average PM2.5 data was collected from 73 fixed air quality monitoring stations that belonged to the Taiwan EPA on the main island of Taiwan. Nearly 280,000 observations from 2006 to 2016 were used for the analysis. Several datasets were collected to determine spatial predictor variables, including the EPA environmental resources dataset, a meteorological dataset, a land-use inventory, a landmark dataset, a digital road network map, a digital terrain model, MODIS Normalized Difference Vegetation Index (NDVI) database, and a power plant distribution dataset. First, conventional LUR and Hybrid Kriging-LUR were utilized to identify the important predictor variables. Then, deep neural network, random forest, and XGBoost algorithms were used to fit the prediction model based on the variables selected by the LUR models. Data splitting, 10-fold cross validation, external data verification, and seasonal-based and county-based validation methods were used to verify the robustness of the developed models. The results demonstrated that the proposed conventional LUR and Hybrid Kriging-LUR models captured 58% and 89% of PM2.5 variations, respectively. When XGBoost algorithm was incorporated, the explanatory power of the models increased to 73% and 94%, respectively. The Hybrid Kriging-LUR with XGBoost algorithm outperformed the other integrated methods. This study demonstrates the value of combining Hybrid Kriging-LUR model and an XGBoost algorithm for estimating the spatial-temporal variability of PM2.5 exposures.
Copyright © 2021 Elsevier Ltd. All rights reserved.

Entities:  

Keywords:  Extreme gradient boosting; Land-use regression; Machine learning; PM(2.5); Variable selection

Mesh:

Substances:

Year:  2021        PMID: 33735646     DOI: 10.1016/j.envpol.2021.116846

Source DB:  PubMed          Journal:  Environ Pollut        ISSN: 0269-7491            Impact factor:   8.071


  1 in total

1.  Development and Evaluation of Spatio-Temporal Air Pollution Exposure Models and Their Combinations in the Greater London Area, UK.

Authors:  Konstantina Dimakopoulou; Evangelia Samoli; Antonis Analitis; Joel Schwartz; Sean Beevers; Nutthida Kitwiroon; Andrew Beddows; Benjamin Barratt; Sophia Rodopoulou; Sofia Zafeiratou; John Gulliver; Klea Katsouyanni
Journal:  Int J Environ Res Public Health       Date:  2022-04-28       Impact factor: 4.614

  1 in total

北京卡尤迪生物科技股份有限公司 © 2022-2023.