Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
3
datasets available to search
ShareScore release 0.9.0
Dataset results
3 results for “Air quality prediction”
The features of the selected papers in the field of air quality prediction
<p>The table is a part of a submitted manuscript (Iskandaryan, D., Ramos, F., & Trilles, S. The Role of Datasets in Air Quality Prediction. Submitted to Atmosphere.) and includes the following features extracted from the selected papers: <em>Year, Case Study, Prediction Target, Dataset Type, Data Rate, Period (Days), Open Data, Algorithm, Time Granularity and Evaluation Metric</em>. The relevant papers were selected from a systematic review in <em>Air Quality Prediction Using Machine Learning Technologies. </em>The works were queried in Association for Computing Machinery, IEEE Xplore, Scopus and Web of Science databases using the following query: ("machine learning") AND ("prediction"OR "forecast") AND ("air quality" OR "air pollution"), which was being applied to title, abstract and keywords. After filtering the results guided by the Preferred Reporting Items for Systematic Reviews and Meta-Analyses, ninety-three papers were selected. The goal of this review is to understand which features are used in the field, in particular to answer the following questions: 1) What types of datasets are used to improve air quality predictions?; and 2) What characteristics of the dataset are important for efficient and effective air quality forecasting? <br> Twenty-six datasets were used by the authors as supplemental air quality data in order to predict air quality more accurately. Those datasets are: "MET"- meteorological data; "Spatial"- topographical characteristics, the locations of the stations; "Temporal"-includes the day of the month, day of the week, the hour of the day; "AOD"- aerosol optical depth; "Social Media"- microblog data; "Traffic"; "PBL Height"- planetary boundary layer height; "Land Use"; "BEV"- Built Environment Variables; "UV Index"; "SP"- Sound Pressure; "PD"-Population Density; "Human Movements"- floating population and estimated traffic volume; "Altitude"; "OMI-SO2"-Satellite-retrieved SO2 from Ozone Monitoring Instrument-SO2; "PPS"- Pollution Point Source; "TS"-Transportation Source; "WFD’"- weather forecast data; "POI Distribution"; "FAPE"- factory air pollution emission; "RND"- Road Network Distribution; "Elevation"; "AEI"- Anthropogenic Emission Inventory; "NDVI"; "Chemical"- chemical component forecast data (organic carbon, black carbon, sea salt, etc.); "Emission".</p>
Air quality data from the article "Typhoon-associated air quality over the Guangdong–Hong Kong–Macao Greater Bay Area, China: machine-learning-based prediction and assessment"
<p>This dataset consists of 26 files. The descriptions of the files are as follows:</p> <ul> <li>aqi_TY.csv, pm25_TY.csv, pm10_TY.csv, so2_TY.csv, no2_TY.csv and o3_TY.csv are the observed values of AQI and concentrations of PM<sub>2.5</sub>, PM<sub>10</sub>, SO<sub>2</sub>, NO<sub>2</sub> and O<sub>3</sub> of 36 monitoring stations used in model establish stage on TY days. The time range is June 2014 to December 2020.</li> <li>aqi_NTY.csv, pm25_NTY.csv, pm10_NTY.csv, so2_NTY.csv, no2_NTY.csv and o3_NTY.csv are the observed values of AQI and concentrations of PM<sub>2.5</sub>, PM<sub>10</sub>, SO<sub>2</sub>, NO<sub>2</sub> and O<sub>3</sub> of 36 monitoring stations used in model establish stage on NTY days. The time range is June 2014 to December 2020.</li> <li>station_info.csv is the detailed information of the 36 monitoring stations used in model establish stage, including station number, city, longitude and latitude.</li> <li>aqi_TY_testing.csv, pm25_TY_testing.csv, pm10_TY_testing.csv, so2_TY_testing.csv, no2_TY_testing.csv and o3_TY_testing.csv are the observed values of AQI and concentrations of PM<sub>2.5</sub>, PM<sub>10</sub>, SO<sub>2</sub>, NO<sub>2</sub> and O<sub>3</sub> of 3 monitoring stations used for testing the model on TY days. The time range is June 2014 to December 2020.</li> <li>aqi_NTY_testing.csv, pm25_NTY_testing.csv, pm10_NTY_testing.csv, so2_NTY_testing.csv, no2_NTY_testing.csv and o3_NTY_testing.csv are the observed values of AQI and concentrations of PM<sub>2.5</sub>, PM<sub>10</sub>, SO<sub>2</sub>, NO<sub>2</sub> and O<sub>3</sub> of 3 monitoring stations used for testing the model on NTY days. The time range is June 2014 to December 2020.</li> <li>sta_testing.csv is the detailed information of the 3 monitoring stations used for testing the model, including station number, city, longitude and latitude.</li> </ul>
Supplementary Materials for 'Spatiotemporal Prediction of Air Quality Using Machine Learning Techniques'
<p>This package includes supplementary materials used to implement air quality prediction in the city of Madrid. It consists of two main subdirectories: Data and Code. The Data directory contains Raw-Data (air quality, meteorological and traffic data from the period of January-June 2019 and January-June 2020, and the location of air quality and meteorological monitoring stations and traffic measurement points of the city of Madrid) and Processed-Data (the output after raws data has gone through the workflow to meet the requirements corresponding to the implementation of the proposed forecasting approaches). The Code directory contains Process Raw Data, Chapter4-ConvLSTM, Chapter5-BiConvLSTM, and Chapter6-A3T_GCN, which provides the procedure for constructing and implementing the proposed approaches.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.