Abstract
Air pollution is one of the major environmental problems in the industrial and populated cities. Predictive mapping of urban air pollution and sharing the generated maps with the public and city officials have positive impacts on society and environment. This article presents a solution based on distributed processing concepts to generate predictive map of air pollution for the next 24 hours. Apache Hadoop has been utilized as the underlying framework to form a cluster of processing machines. In order to improve the processing speed along with required machine learning functionalities, Apache Spark has been employed on the Hadoop cluster. The solution enables us to efficiently predict air quality classes on monitoring stations of Tehran, the capital of Iran for the next 24 hours. Using Inverse distance weighting (IDW) method, the predictive map of air quality classes is generated afterward for the whole city. The results showed that the proposed approach can achieve a reasonable speed in processing of big spatial data along with horizontal scalability..
| Original language | English |
|---|---|
| Title of host publication | 2017 International Conference on Cloud and Big Data Computing, ICCBDC 2017 |
| Publisher | Association for Computing Machinery |
| Pages | 89-93 |
| Number of pages | 5 |
| ISBN (Electronic) | 9781450353434 |
| DOIs | |
| Publication status | Published - 17 Sept 2017 |
| Externally published | Yes |
| Event | International Conference on Cloud and Big Data Computing, ICCBDC 2017 - London, United Kingdom Duration: 17 Sept 2017 → 19 Sept 2017 |
Publication series
| Name | ACM International Conference Proceeding Series |
|---|
Conference
| Conference | International Conference on Cloud and Big Data Computing, ICCBDC 2017 |
|---|---|
| Abbreviated title | ICCBDC 2017 |
| Country/Territory | United Kingdom |
| City | London |
| Period | 17/09/17 → 19/09/17 |
UN SDGs
This output contributes to the following UN Sustainable Development Goals (SDGs)
-
SDG 11 Sustainable Cities and Communities
Keywords
- Air pollution
- Big spatial data
- Distributed processing
- Hadoop
- Predictive mapping
- Spark
- ITC-CV
Fingerprint
Dive into the research topics of 'Predictive mapping of urban air pollution using apache spark on a hadoop cluster'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver