首页    期刊浏览 2024年09月21日 星期六
登录注册

文章基本信息

  • 标题:Data mining methods for prediction of air pollution
  • 本地全文:下载
  • 作者:Krzysztof Siwek ; Stanisław Osowski
  • 期刊名称:International Journal of Applied Mathematics and Computer Science
  • 电子版ISSN:2083-8492
  • 出版年度:2016
  • 卷号:26
  • 期号:2
  • DOI:10.1515/amcs-2016-0033
  • 出版社:De Gruyter Open
  • 摘要:The paper discusses methods of data mining for prediction of air pollution. Two tasks in such a problem are important: generation and selection of the prognostic features, and the final prognostic system of the pollution for the next day. An advanced set of features, created on the basis of the atmospheric parameters, is proposed. This set is subject to analysis and selection of the most important features from the prediction point of view. Two methods of feature selection are compared. One applies a genetic algorithm (a global approach), and the other—a linear method of stepwise fit (a locally optimized approach). On the basis of such analysis, two sets of the most predictive features are selected. These sets take part in prediction of the atmospheric pollutants PM10, SO2, NO2 and O3. Two approaches to prediction are compared. In the first one, the features selected are directly applied to the random forest (RF), which forms an ensemble of decision trees. In the second case, intermediate predictors built on the basis of neural networks (the multilayer perceptron, the radial basis function and the support vector machine) are used. They create an ensemble integrated into the final prognosis. The paper shows that preselection of the most important features, cooperating with an ensemble of predictors, allows increasing the forecasting accuracy of atmospheric pollution in a significant way
  • 关键词:computational intelligence; feature selection; neural networks; random forest; air pollution forecasting
国家哲学社会科学文献中心版权所有