Determination of Trace Organic Contaminant Concentration via Machine Classification of Surface-Enhanced Raman Spectra (2402.00197v1)
Abstract: Accurate detection and analysis of traces of persistent organic pollutants in water is important in many areas, including environmental monitoring and food quality control, due to their long environmental stability and potential bioaccumulation. While conventional analysis of organic pollutants requires expensive equipment, surface enhanced Raman spectroscopy (SERS) has demonstrated great potential for accurate detection of these contaminants. However, SERS analytical difficulties, such as spectral preprocessing, denoising, and substrate-based spectral variation, have hindered widespread use of the technique. Here, we demonstrate an approach for predicting the concentration of sample pollutants from messy, unprocessed Raman data using machine learning. Frequency domain transform methods, including the Fourier and Walsh Hadamard transforms, are applied to sets of Raman spectra of three model micropollutants in water (rhodamine 6G, chlorpyrifos, and triclosan), which are then used to train machine learning algorithms. Using standard machine learning models, the concentration of sample pollutants are predicted with more than 80 percent cross-validation accuracy from raw Raman data. cross-validation accuracy of 85 percent was achieved using deep learning for a moderately sized dataset (100 spectra), and 70 to 80 percent cross-validation accuracy was achieved even for very small datasets (50 spectra). Additionally, standard models were shown to accurately identify characteristic peaks via analysis of their importance scores. The approach shown here has the potential to be applied to facilitate accurate detection and analysis of persistent organic pollutants by surface-enhanced Raman spectroscopy.
- Mackay, D.; Fraser, A. Bioaccumulation of persistent organic chemicals: mechanisms and models. Environmental Pollution 2000, 375–391
- Dafouz, R.; Cáceres, N.; andNicola Mastroianni, J. L. R.-G.; de Alda, M. L.; Barceló, D.; Ángel Gilde Miguel; Valcárcel, Y. Does the presence of caffeine in the marine environment represent an environmental risk? A regional and global study. Science of The Total Environment 2018, 632–642
- Compendium of Canada’s engagement in international environmental agreements and instruments; Environment and Climate Change Canada: 867 Lakeshore Rd Burlington ON L7S 1A1, 2001
- Rozati, R.; Reddy, P.; Reddanna, P.; Mujtaba, R. Role of environmental estrogens in the deterioration of male factor fertility. Fertility and Sterility 2002, 1187–1194
- K.C.Jones; Voogt, P. Persistent organic pollutants (POPs): state of the science. Environmental Pollution 1999,
- Bodelón, G.; Pastoriza-Santos, I. Recent progress in Surface-Enhanced Raman Scattering for the detection of chemical contaminants in water. Frontiers in Chemistry 2020,
- Lussier, F.; Thibault, V.; Charron, B.; Q.Wallace, G.; Masson, J.-F. Deep learning and artificial intelligence methods for Raman and surface-enhanced Raman scattering. Trends in Analytical Chemistry 2020,
- Dabodiya, T. S.; Sontti, S. G.; Wei, Z.; Lu, Q.; Billet, R.; Murugan, A. V.; Zhang, X. Ultrasensitive Surface-Enhanced Raman Spectroscopy detection by porous silver supraparticles from self–lubricating drop evaporation. Advanced Materials Interfaces 2022,
- Jones, R. R.; Hooper, D. C.; Zhang, L.; Wolverson, D.; Valev, V. K. Raman techniques: Fundamentals and frontiers. Nanoscale Research Letters 2019,
- Wu, M.; Wang, S.; Pan, S.; Terentis, A. C.; Strasswimmer, J.; Zhu, X. Deep learning data augmentation for Raman spectroscopy cancer tissue classification. Scientific Reports 2021,
- Kanike, C.; Wu, H.; W., Z. A.; Li, Y.; Wei, Z.; Unsworth, L. D.; Atta, A.; Zhang, X. Flow-based approach for scalable fabrication of Ag nanostructured substrate as a platform for surface-enhanced Raman scattering. Manuscript to be Submitted 2023,
- Maruthamuthu, M. K.; Raffiee, A. H.; Oliveira, D. M. D.; Ardekani, A. M.; Verma, M. S. Raman spectra-based deep learning: A tool to identify microbial contamination. Microbiology Open 2020,
- Fan, X.; Ming, W.; Zeng, H.; Zhang, Z.; Lu, H. Deep learning-based component identification for the Raman spectra of mixtures. Analyst 2019, 1789–1798
- Marshall, A. G.; Comisarow, M. B. Fourier and Hadamard transform methods in spectroscopy. Analytical Chemistry 1975, 491–504
- Majumdar, S.; Laha, A. K. Clustering and classification of time series using topological data analysis with applications to finance. Expert Systems With Applications 2020, 162
- Faouzi, J. Machine Learning (Emerging Trends and Applications); ProudPen: Paris, France
- Brownlee, J. 1D convolutional neural network models for human activity recognition. 2020; https://machinelearningmastery.com/cnn-models-for-human-activity-recognition-time-series-classification/
- Fawaz, H. I.; Forestier, G.; Weber, J.; Idoumghar, L.; Muller, P.-A. Deep learning for time series classification: a review. Data Mining and Knowledge Discovery 2019, 917–963
- Loken, C. et al. SciNet: Lessons Learned from Building a Power-efficient Top-20 System and Data Centre. Journal of Physics: Conference Series 2010, 256
- Svensson, O.; Josefson, M.; W.Langkilde, F. Reaction monitoring using Raman spectroscopy and chemometrics. Chemometrics and Intelligent Laboratory Systems 1999, 49