Using remotely sensed data for air pollution assessment (2402.06653v1)
Abstract: Air pollution constitutes a global problem of paramount importance that affects not only human health, but also the environment. The existence of spatial and temporal data regarding the concentrations of pollutants is crucial for performing air pollution studies and monitor emissions. However, although observation data presents great temporal coverage, the number of stations is very limited and they are usually built in more populated areas. The main objective of this work is to create models capable of inferring pollutant concentrations in locations where no observation data exists. A machine learning model, more specifically the random forest model, was developed for predicting concentrations in the Iberian Peninsula in 2019 for five selected pollutants: $NO_2$, $O_3$ $SO_2$, $PM10$, and $PM2.5$. Model features include satellite measurements, meteorological variables, land use classification, temporal variables (month, day of year), and spatial variables (latitude, longitude, altitude). The models were evaluated using various methods, including station 10-fold cross-validation, in which in each fold observations from 10\% of the stations are used as testing data and the rest as training data. The $R2$, RMSE and mean bias were determined for each model. The $NO_2$ and $O_3$ models presented good values of $R2$, 0.5524 and 0.7462, respectively. However, the $SO_2$, $PM10$, and $PM2.5$ models performed very poorly in this regard, with $R2$ values of -0.0231, 0.3722, and 0.3303, respectively. All models slightly overestimated the ground concentrations, except the $O_3$ model. All models presented acceptable cross-validation RMSE, except the $O_3$ and $PM10$ models where the mean value was a little higher (12.5934 $\mu g/m3$ and 10.4737 $\mu g/m3$, respectively).
- D. W. Dockery, C. A. Pope, X. Xu, J. D. Spengler, J. H. Ware, M. E. Fay, B. G. Ferris Jr, and F. E. Speizer, “An association between air pollution and mortality in six us cities,” New England journal of medicine, vol. 329, no. 24, pp. 1753–1759, 1993.
- W. H. Organization, “Air quality and health,” (Accessed May 2022). [Online]. Available: https://www.who.int/teams/environment-climate-change-and-health/air-quality-and-health/health-impacts
- E. E. Agency, “Air quality in europe - 2020 report eea, report no 9/2020,” 2020.
- P. Veefkind, R. Van Oss, H. Eskes, A. Borowiak, F. Dentner, and J. Wilson, “The applicability of remote sensing in the field of air pollution,” Institute for Environment and Sustainability, Italy, vol. 59, 2007.
- E. E. Agency, “European air quality portal,” (Accessed February 2022). [Online]. Available: https://aqportal.discomap.eea.europa.eu/
- E. S. Agency, “Sentinel-5p pre-operations data hub,” (Accessed February 2022). [Online]. Available: https://s5phub.copernicus.eu/dhus/#/home
- “Era5 hourly data on single levels from 1979 to present,” (Accessed February 2022). [Online]. Available: https://cds.climate.copernicus.eu/cdsapp#!/dataset/reanalysis-era5-single-levels?tab=overview
- E. E. Agency, “Corine land cover,” (Accessed February 2022). [Online]. Available: https://land.copernicus.eu/pan-european/corine-land-cover
- ——, “Air pollution sources,” (Accessed May 2022). [Online]. Available: https://www.eea.europa.eu/themes/air/air-pollution-sources-1
- J. Geiger, L. Malherbe, F. Mathe, M. Ross-Jones, K. Sjoberg, W. Spangl, B. Stacey, A. Ortiz, F. de Leeuw, A. Borowiak et al., “Assessment on siting criteria, classification and representativeness of air quality monitoring stations. jrc–aquila position paper, 2013,” 2013.
- NASA, “Remote sensing: An overview,” (Accessed May 2022). [Online]. Available: https://earthdata.nasa.gov/learn/backgrounders/remote-sensing
- J. A. Engel-Cox, R. M. Hoff, and A. Haymet, “Recommendations on the use of satellite remote-sensing data for urban air quality,” Journal of the Air & Waste Management Association, vol. 54, no. 11, pp. 1360–1371, 2004.
- S. A. Christopher and P. Gupta, “Satellite remote sensing of particulate matter air quality: The cloud-cover problem,” Journal of the Air & Waste Management Association, vol. 60, no. 5, pp. 596–602, 2010.
- M. Mirzaei, S. Bertazzon, I. Couloigner, B. Farjad, and R. Ngom, “Estimation of local daily pm2. 5 concentration during wildfire episodes: integrating modis aod with multivariate linear mixed effect (lme) models,” Air Quality, Atmosphere & Health, vol. 13, no. 2, pp. 173–185, 2020.
- K. Qin, L. Rao, J. Xu, Y. Bai, J. Zou, N. Hao, S. Li, and C. Yu, “Estimating ground level no2 concentrations over central-eastern china using a satellite-based geographically and temporally weighted regression model,” Remote Sensing, vol. 9, no. 9, p. 950, 2017.
- H. J. Lee, B. A. Coull, M. L. Bell, and P. Koutrakis, “Use of satellite-based aerosol optical depth and spatial clustering to predict ambient pm2. 5 concentrations,” Environmental research, vol. 118, pp. 8–15, 2012.
- J. Chen, K. de Hoogh, J. Gulliver, B. Hoffmann, O. Hertel, M. Ketzel, M. Bauwelinck, A. Van Donkelaar, U. A. Hvidtfeldt, K. Katsouyanni et al., “A comparison of linear regression, regularization, and machine learning algorithms to develop europe-wide spatial models of fine particles and nitrogen dioxide,” Environment international, vol. 130, p. 104934, 2019.
- G. Chen, S. Li, L. D. Knibbs, N. A. Hamm, W. Cao, T. Li, J. Guo, H. Ren, M. J. Abramson, and Y. Guo, “A machine learning method to estimate pm2. 5 concentrations across china with remote sensing, meteorological and land use information,” Science of the Total Environment, vol. 636, pp. 52–60, 2018.
- B. Mahesh, “Machine learning algorithms-a review,” International Journal of Science and Research (IJSR).[Internet], vol. 9, pp. 381–386, 2020.
- C.-M. Vong, W.-F. Ip, P.-k. Wong, and J.-y. Yang, “Short-term prediction of air pollution in macau using support vector machines,” Journal of Control Science and Engineering, vol. 2012, 2012.
- A. S. Sánchez, P. G. Nieto, P. R. Fernández, J. del Coz Díaz, and F. J. Iglesias-Rodríguez, “Application of an svm-based regression model to the air quality study at local scale in the avilés urban area (spain),” Mathematical and Computer Modelling, vol. 54, no. 5-6, pp. 1453–1466, 2011.
- A. Alimissis, K. Philippopoulos, C. Tzanis, and D. Deligiorgi, “Spatial estimation of urban air pollution with the use of artificial neural network models,” Atmospheric environment, vol. 191, pp. 205–213, 2018.
- R. Schneider, A. M. Vicedo-Cabrera, F. Sera, P. Masselot, M. Stafoggia, K. de Hoogh, I. Kloog, S. Reis, M. Vieno, and A. Gasparrini, “A satellite-based spatio-temporal machine learning model to reconstruct daily pm2. 5 concentrations across great britain,” Remote sensing, vol. 12, no. 22, p. 3803, 2020.
- E. S. Agency, “Copernicus sentinel-5p,” (Accessed February 2022). [Online]. Available: https://sentinels.copernicus.eu/web/sentinel/missions/sentinel-5p
- H. Omrani, B. Omrani, B. Parmentier, and M. Helbich, “S5p-tools,” (Accessed February 2022). [Online]. Available: https://github.com/bilelomrani1/s5p-tools
- H. Eskes, J. V. Geffen, F. Boersma, K.-U. Eichmann, A. Apituley, M. Pedergnana, M. Sneep, J. P. Veefkind, and D. Loyola, “Sentinel-5 precursor/tropomi level 2 product user manual nitrogendioxide document number : S5p-knmi-l2-0021-ma.”
- “Atmospheric toolbox - harp,” (Accessed February 2022). [Online]. Available: https://atmospherictoolbox.org/harp/
- “Climate data store,” (Accessed February 2022). [Online]. Available: https://cds.climate.copernicus.eu/cdsapp#!/home
- S. learn documentation, “scikit learn - machine learning in python,” (Accessed July 2022). [Online]. Available: https://scikit-learn.org/stable/index.html
- ECMWF, “Era5 data documentation,” (Accessed February 2022). [Online]. Available: https://confluence.ecmwf.int/display/CKB/ERA5
- P. S. Monks, A. Archibald, A. Colette, O. Cooper, M. Coyle, R. Derwent, D. Fowler, C. Granier, K. S. Law, G. Mills et al., “Tropospheric ozone and its precursors from the urban to the global scale from air quality to short-lived climate forcer,” Atmospheric Chemistry and Physics, vol. 15, no. 15, pp. 8889–8973, 2015.
- E. S. Agency, “Level-2 algorithms - aerosol index,” (Accessed October 2022). [Online]. Available: https://sentinels.copernicus.eu/web/sentinel/technical-guides/sentinel-5p/level-2/aerosol-index