Difference between revisions of "IS428 2018-19 T1 Assign Fu Yu"
Yu.fu.2015 (talk | contribs) |
Yu.fu.2015 (talk | contribs) |
||
Line 67: | Line 67: | ||
[[File:A typical day in sofia city yu.fu.2015.png|center]] | [[File:A typical day in sofia city yu.fu.2015.png|center]] | ||
− | <p>To investigate how PM10 concentrations vary from hours to hours in a day, one can simply highlight a day on the heatmap and the hourly PM10 concentration graphs will update accordingly. For example, as shown in the diagram above, on 8 January 2018, air quality in stations STA-BG0052A and STA-BG0050A is poor after 10pm while air quality in station STA-BG0073A is poor starting from 12pm and in STA-BG0040A, air quality is poor throughout the | + | <p>To investigate how PM10 concentrations vary from hours to hours in a day, one can simply highlight a day on the heatmap and the hourly PM10 concentration graphs will update accordingly. For example, as shown in the diagram above, on 8 January 2018, air quality in stations STA-BG0052A and STA-BG0050A is poor after 10pm while air quality in station STA-BG0073A is poor starting from 12pm and in STA-BG0040A, air quality is poor throughout the day.</p> |
+ | |||
+ | <p> <b>Trends:</b> | ||
+ | * There is no obvious trend showing on what time of a day the PM10 concentration is higher or lower | ||
+ | * Nevertheless, PM10 concentrations do vary from time to time in a day. | ||
− | |||
− | |||
+ | |} | ||
+ | {| class="wikitable" style="background-color:#FFFFFF;" width="100%" | ||
+ | |'''1.3 Anomalies in the Datesetss | ||
+ | ''' | ||
+ | * From 2013 to 2015, hourly PM10 concentration data is not available. Data was either only collected at 00:00 once or the hour when the data was collected was not recorded from 2013 to 2015. Also, in 2016, data was only collected at 00:00 on some days. If the data was only collected at 00:00 or only collected once, it might affect the accuracy of PM10 concentration distribution across the year because PM10 concentrations are different at different time in a day. It might happen that at the time the data was collected, the PM10 concentration was too high or too low, which would not be representative of the PM10 concentration of a day. | ||
+ | * For 2017, only November and December data is available. Hence changes in air quality from 2016 to 2017 could not be investigated which might give the insights of why air quality in 2018 has improved. | ||
|} | |} | ||
Revision as of 00:41, 11 November 2018
Contents
Problem & Motivation
Air pollution is an important risk factor for health in Europe and worldwide. A recent review of the global burden of disease showed that it is one of the top ten risk factors for health globally. Worldwide an estimated 7 million people died prematurely because of pollution; in the European Union (EU) 400,000 people suffer a premature death. The Organisation for Economic Cooperation and Development (OECD) predicts that in 2050 outdoor air pollution will be the top cause of environmentally related deaths worldwide. In addition, air pollution has also been classified as the leading environmental cause of cancer.
Air quality in Bulgaria is a big concern: measurements show that citizens all over the country breathe in air that is considered harmful to health. For example, concentrations of PM2.5 and PM10 are much higher than what the EU and the World Health Organization (WHO) have set to protect health.
Bulgaria had the highest PM2.5 concentrations of all EU-28 member states in urban areas over a three-year average. For PM10, Bulgaria is also leading on the top polluted countries with 77 μg/m3on the daily mean concentration (EU limit value is 50 μg/m3).
According to the WHO, 60 percent of the urban population in Bulgaria is exposed to dangerous (unhealthy) levels of particulate matter (PM10).
Dataset Analysis & Transformation Process
Decode the geohash column in Air Tube data files
Geohash tells the station locations. However Tableau is not able to interpret geohash as geographic data. Before Air Tube data is imported to Tableau for analysis, geohash needs to be decoded into geographical coordinates. As the two Air Tube data files- data_bg_2017.xlsx and data_bg_2018.xlsx are of big sizes and there are duplicate geohash records in the data, an Excel file containing a unique geohash list was created.
Step 1: Use "pygeohash" package to decode the geohash list and output the coordinates in an Excel file
Step 2: Combine geohash list and coordinates list into one Excel file and update the coordinates, latitude, longitude in data_bg_2017.xlsx and data_bg_2018.xlsx using VLOOKUP, LEFT and RIGHT functions in Excel
Dataset Import Structure & Process
Interactive Visualization
Interesting & Anomalous Observations
Task 1: Spatio-temporal Analysis of Official Air Quality
1.2 A Typical Day in Sofia City
To investigate how PM10 concentrations vary from hours to hours in a day, one can simply highlight a day on the heatmap and the hourly PM10 concentration graphs will update accordingly. For example, as shown in the diagram above, on 8 January 2018, air quality in stations STA-BG0052A and STA-BG0050A is poor after 10pm while air quality in station STA-BG0073A is poor starting from 12pm and in STA-BG0040A, air quality is poor throughout the day. Trends:
|
1.3 Anomalies in the Datesetss
|