使用机器学习技术的空气质量分析和PM2.5建模:对印度海德拉巴市的研究

IF 2.1 Q2 ENGINEERING, MULTIDISCIPLINARY Cogent Engineering Pub Date : 2023-08-13 DOI:10.1080/23311916.2023.2243743
Aneesh Mathew, P. R. Gokul, Padala Raja Shekar, K. S. Arunab, Hazem Ghassan Abdo, Hussein Almohamad, Ahmed Abdullah Al Dughairi
{"title":"使用机器学习技术的空气质量分析和PM2.5建模:对印度海德拉巴市的研究","authors":"Aneesh Mathew, P. R. Gokul, Padala Raja Shekar, K. S. Arunab, Hazem Ghassan Abdo, Hussein Almohamad, Ahmed Abdullah Al Dughairi","doi":"10.1080/23311916.2023.2243743","DOIUrl":null,"url":null,"abstract":"Abstract The rapid urbanization and industrialization in many parts of the world have made air pollution a global public health problem. A study conducted by the Swiss organization IQAir indicated that 22 of the top 30 most polluted cities in the world are in India. This creates the problem of air pollution, which is very relevant to India as well. Exposure to air pollutants has both acute (short-term) and chronic (long-term) impacts on health. Among the major air pollutants, particulate matter 2.5 (PM2.5) is the most harmful, and its long-term exposure can impair lung functions. Pollutant concentrations vary temporally and are dependent on the local meteorology and emissions at a given geographic location. PM2.5 forecasting models have the potential to develop strategies for evaluating and alerting the public regarding expected hazardous levels of air pollution. Accurate measurement and forecasting of pollutant concentrations are critical for assessing air quality and making informed strategic decisions. Recently, data-driven machine learning algorithms for PM2.5 forecasting have received a lot of attention. In this work, a spatio-temporal analysis of air quality was first performed for Hyderabad, indicating that average PM2.5 concentrations during the winter were 68% higher than those during the summer. Following that, PM2.5 modelling was done using three different techniques: multilinear regression, K-nearest neighbours (KNN), and histogram-based gradient boost (HGBoost). Among these, the HGBoost regression model, which used both pollution and meteorological data as inputs, outperformed the other two techniques. During testing, the model acquired an amazing R2 value of 0.859, suggesting a significant connection with the actual data. Additionally, the model exhibited a minimum Mean Absolute Error (MAE) of 5.717 μg/m3 and a Root Mean Square Error (RMSE) of 7.647 μg/m3, further confirming its accuracy in predicting PM2.5 concentrations. In our investigation, we discovered that the HGBoost3 model beat other PM2.5 modelling models by having the lowest error and the highest R2 value. This study made a substantial addition by incorporating the spatiotemporal relationship between air pollutants and meteorological variables in predicting air quality. This method has the potential to improve the creation of more precise air pollution forecast models.","PeriodicalId":10464,"journal":{"name":"Cogent Engineering","volume":" ","pages":""},"PeriodicalIF":2.1000,"publicationDate":"2023-08-13","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"0","resultStr":"{\"title\":\"Air quality analysis and PM2.5 modelling using machine learning techniques: A study of Hyderabad city in India\",\"authors\":\"Aneesh Mathew, P. R. Gokul, Padala Raja Shekar, K. S. Arunab, Hazem Ghassan Abdo, Hussein Almohamad, Ahmed Abdullah Al Dughairi\",\"doi\":\"10.1080/23311916.2023.2243743\",\"DOIUrl\":null,\"url\":null,\"abstract\":\"Abstract The rapid urbanization and industrialization in many parts of the world have made air pollution a global public health problem. A study conducted by the Swiss organization IQAir indicated that 22 of the top 30 most polluted cities in the world are in India. This creates the problem of air pollution, which is very relevant to India as well. Exposure to air pollutants has both acute (short-term) and chronic (long-term) impacts on health. Among the major air pollutants, particulate matter 2.5 (PM2.5) is the most harmful, and its long-term exposure can impair lung functions. Pollutant concentrations vary temporally and are dependent on the local meteorology and emissions at a given geographic location. PM2.5 forecasting models have the potential to develop strategies for evaluating and alerting the public regarding expected hazardous levels of air pollution. Accurate measurement and forecasting of pollutant concentrations are critical for assessing air quality and making informed strategic decisions. Recently, data-driven machine learning algorithms for PM2.5 forecasting have received a lot of attention. In this work, a spatio-temporal analysis of air quality was first performed for Hyderabad, indicating that average PM2.5 concentrations during the winter were 68% higher than those during the summer. Following that, PM2.5 modelling was done using three different techniques: multilinear regression, K-nearest neighbours (KNN), and histogram-based gradient boost (HGBoost). Among these, the HGBoost regression model, which used both pollution and meteorological data as inputs, outperformed the other two techniques. During testing, the model acquired an amazing R2 value of 0.859, suggesting a significant connection with the actual data. Additionally, the model exhibited a minimum Mean Absolute Error (MAE) of 5.717 μg/m3 and a Root Mean Square Error (RMSE) of 7.647 μg/m3, further confirming its accuracy in predicting PM2.5 concentrations. In our investigation, we discovered that the HGBoost3 model beat other PM2.5 modelling models by having the lowest error and the highest R2 value. This study made a substantial addition by incorporating the spatiotemporal relationship between air pollutants and meteorological variables in predicting air quality. This method has the potential to improve the creation of more precise air pollution forecast models.\",\"PeriodicalId\":10464,\"journal\":{\"name\":\"Cogent Engineering\",\"volume\":\" \",\"pages\":\"\"},\"PeriodicalIF\":2.1000,\"publicationDate\":\"2023-08-13\",\"publicationTypes\":\"Journal Article\",\"fieldsOfStudy\":null,\"isOpenAccess\":false,\"openAccessPdf\":\"\",\"citationCount\":\"0\",\"resultStr\":null,\"platform\":\"Semanticscholar\",\"paperid\":null,\"PeriodicalName\":\"Cogent Engineering\",\"FirstCategoryId\":\"1085\",\"ListUrlMain\":\"https://doi.org/10.1080/23311916.2023.2243743\",\"RegionNum\":0,\"RegionCategory\":null,\"ArticlePicture\":[],\"TitleCN\":null,\"AbstractTextCN\":null,\"PMCID\":null,\"EPubDate\":\"\",\"PubModel\":\"\",\"JCR\":\"Q2\",\"JCRName\":\"ENGINEERING, MULTIDISCIPLINARY\",\"Score\":null,\"Total\":0}","platform":"Semanticscholar","paperid":null,"PeriodicalName":"Cogent Engineering","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1080/23311916.2023.2243743","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"Q2","JCRName":"ENGINEERING, MULTIDISCIPLINARY","Score":null,"Total":0}
引用次数: 0
查看原文
分享 分享
微信好友 朋友圈 QQ好友 复制链接
本刊更多论文
Air quality analysis and PM2.5 modelling using machine learning techniques: A study of Hyderabad city in India
Abstract The rapid urbanization and industrialization in many parts of the world have made air pollution a global public health problem. A study conducted by the Swiss organization IQAir indicated that 22 of the top 30 most polluted cities in the world are in India. This creates the problem of air pollution, which is very relevant to India as well. Exposure to air pollutants has both acute (short-term) and chronic (long-term) impacts on health. Among the major air pollutants, particulate matter 2.5 (PM2.5) is the most harmful, and its long-term exposure can impair lung functions. Pollutant concentrations vary temporally and are dependent on the local meteorology and emissions at a given geographic location. PM2.5 forecasting models have the potential to develop strategies for evaluating and alerting the public regarding expected hazardous levels of air pollution. Accurate measurement and forecasting of pollutant concentrations are critical for assessing air quality and making informed strategic decisions. Recently, data-driven machine learning algorithms for PM2.5 forecasting have received a lot of attention. In this work, a spatio-temporal analysis of air quality was first performed for Hyderabad, indicating that average PM2.5 concentrations during the winter were 68% higher than those during the summer. Following that, PM2.5 modelling was done using three different techniques: multilinear regression, K-nearest neighbours (KNN), and histogram-based gradient boost (HGBoost). Among these, the HGBoost regression model, which used both pollution and meteorological data as inputs, outperformed the other two techniques. During testing, the model acquired an amazing R2 value of 0.859, suggesting a significant connection with the actual data. Additionally, the model exhibited a minimum Mean Absolute Error (MAE) of 5.717 μg/m3 and a Root Mean Square Error (RMSE) of 7.647 μg/m3, further confirming its accuracy in predicting PM2.5 concentrations. In our investigation, we discovered that the HGBoost3 model beat other PM2.5 modelling models by having the lowest error and the highest R2 value. This study made a substantial addition by incorporating the spatiotemporal relationship between air pollutants and meteorological variables in predicting air quality. This method has the potential to improve the creation of more precise air pollution forecast models.
求助全文
通过发布文献求助,成功后即可免费获取论文全文。 去求助
来源期刊
Cogent Engineering
Cogent Engineering ENGINEERING, MULTIDISCIPLINARY-
CiteScore
4.00
自引率
5.30%
发文量
213
审稿时长
13 weeks
期刊介绍: One of the largest, multidisciplinary open access engineering journals of peer-reviewed research, Cogent Engineering, part of the Taylor & Francis Group, covers all areas of engineering and technology, from chemical engineering to computer science, and mechanical to materials engineering. Cogent Engineering encourages interdisciplinary research and also accepts negative results, software article, replication studies and reviews.
期刊最新文献
Evaluating road work site safety management: A case study of the Amman bus rapid transit project construction Toward optimizing scientific workflow using multi-objective optimization in a cloud environment Technology capability of Indonesian medium-sized shipyards for ship production using Product-oriented Work Breakdown Structure method (case study on shipbuilding of Mini LNG vessel) NREL Phase VI wind turbine blade tip with S809 airfoil profile winglet design and performance analysis using computational fluid dynamics Revisiting multi-domain empirical modelling of light-emitting diode luminaire
×
引用
GB/T 7714-2015
复制
MLA
复制
APA
复制
导出至
BibTeX EndNote RefMan NoteFirst NoteExpress
×
×
提示
您的信息不完整,为了账户安全,请先补充。
现在去补充
×
提示
您因"违规操作"
具体请查看互助需知
我知道了
×
提示
现在去查看 取消
×
提示
确定
0
微信
客服QQ
Book学术公众号 扫码关注我们
反馈
×
意见反馈
请填写您的意见或建议
请填写您的手机或邮箱
已复制链接
已复制链接
快去分享给好友吧!
我知道了
×
扫码分享
扫码分享
Book学术官方微信
Book学术文献互助
Book学术文献互助群
群 号:481959085
Book学术
文献互助 智能选刊 最新文献 互助须知 联系我们:info@booksci.cn
Book学术提供免费学术资源搜索服务,方便国内外学者检索中英文文献。致力于提供最便捷和优质的服务体验。
Copyright © 2023 Book学术 All rights reserved.
ghs 京公网安备 11010802042870号 京ICP备2023020795号-1