Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
9
datasets available to search
ShareScore release 0.9.0
Dataset results
9 results for “entso-e”
Energy Climate dataset consitent with ENTSO-E TYNDP2020 studies (CSV & NetCDF) for ACDC-ESM
<p><strong>Energy Climate dataset consistent with ENTSO-E Pan-European Climatic Database (PECD 2021.3) in CSV and netCDF format</strong></p> <p><strong>TL;DR</strong>: this is a tidy and friendly version of a recreation of ENTSO-E's PECD 2021.3 data by using ERA5: hourly capacity factors for wind onshore, offshore, solar PV and hourly electricity demand are provided. All the data is provided for 28-71 climatic years (1950-2020 for wind and solar, 1982-2010 for demand).</p> <p><strong>Description</strong><br> Country averages of energy-climate variables generated using the Python scripts, based on the <a href="https://2020.entsos-tyndp-scenarios.eu/">ENTSO-E's TYNDP 2020 study</a>. For the following scenario's data is available</p> <ul> <li>National trends 2025 (NT 2025)</li> <li>National trends 2030 (NT 2030)</li> <li>National trends 2040 (NT 2040)</li> <li>Distributed Energy 2030 (DE 2030)</li> <li>Distributed Energy 2040 (DE 2040)</li> <li>Global Ambitions (GA 2030)</li> <li>Global Ambitions (GA 2040)</li> </ul> <p>The time-series are at hourly resolution and the included variables are:</p> <ul> <li>Generation wind offshore (aggregated for all years per scenario in a .zip)</li> <li>Generation wind onshore (aggregated for all years per scenario in a .zip)</li> <li>Generation solar photovoltaic (aggregated for all years per scenario in a .zip)</li> <li>Total energy demand (all zones combined in single file per scenario)</li> </ul> <p>The Files are provided in CSV (.csv) & NetCDF (.nc). The data is given per ENTSO-E's bidding zone as used within the TYNDP2020.<br> </p> <p><strong>DISCLAIMER</strong>: <em>the content of this dataset has been created with the greatest possible care. However, we invite to use the original data for critical applications and studies. </em></p>
ENTSO-E Pan-European Climatic Database (PECD 2021.3) in Parquet format
<p><strong>ENTSO-E Pan-European Climatic Database (PECD 2021.3) in Parquet format</strong></p> <p><strong>TL;DR</strong>: this is a tidy and friendly version of a subset of the PECD 2021.3 data by ENTSO-E: hourly capacity factors for wind onshore, offshore, solar PV, hourly electricity demand, weekly inflow for reservoir and pumping and daily generation for run-of-river. All the data is provided for >30 climatic years (1982-2019 for wind and solar, 1982-2016 for demand, 1982-2017 for hydropower) and at national and sub-national (>140 zones) level.</p> <p><strong>UPDATE (19/10/2022): </strong>updated the demand files due after fixing a bug in the processing code (the file for 2030 was the same for 2025) and solving an issue caused by a malformed header in the ENTSO-E excel files. </p> <p> </p> <p> </p> <p>ENTSO-E has released with the latest European Resource Adequacy Assessment (<a href="https://www.entsoe.eu/outlooks/eraa/">ERAA 2021</a>) all the inputs used in the study.<br> Those inputs include:<br> - Demand dataset: <a href="https://eepublicdownloads.azureedge.net/clean-documents/sdc-documents/ERAA/Demand%20Dataset.7z">https://eepublicdownloads.azureedge.net/clean-documents/sdc-documents/ERAA/Demand%20Dataset.7z</a><br> - Climate data: <a href="https://eepublicdownloads.entsoe.eu/clean-documents/sdc-documents/ERAA/Climate%20Data.7z">https://eepublicdownloads.entsoe.eu/clean-documents/sdc-documents/ERAA/Climate%20Data.7z</a></p> <p>The data files and the methodology are available on the <a href="https://www.entsoe.eu/outlooks/eraa/2021/eraa-downloads/">official webpage</a>. </p> <p>As done for the previous releases (see <a href="https://zenodo.org/record/3702418#.YbmhR23MKMo">https://zenodo.org/record/3702418#.YbmhR23MKMo</a> and <a href="https://zenodo.org/record/3985078#.Ybmhem3MKMo">https://zenodo.org/record/3985078#.Ybmhem3MKMo</a>), the original data - stored in large Excel spreadsheets - have been tidied and formatted in open and friendly formats (CSV for the small tables and Parquet for the large files)</p> <p>Furthermore, we have carried out a simple country-aggregation for the original data - that uses instead >140 zones.</p> <p><strong>DISCLAIMER</strong>: <em>the content of this dataset has been created with the greatest possible care. However, we invite to use the original data for critical applications and studies. </em></p> <p><strong>Description</strong></p> <p>This dataset includes the following files:</p> <p>- <em>capacities-national-estimates.csv</em>: installed capacity in MW per zone, technology and the two scenarios (2025 and 2030). The files include also the total capacity for each technology per country (sum of all the zones within a country)<br> - <em>PECD-2021.3-wide-LFSolarPV-2025</em> and <em>PECD-2021.3-wide-LFSolarPV-2030</em>: tables in Parquet format storing in each row the capacity factor for solar PV for a hour of the year and all the climatic years (1982-2019) for a specific zone. The two files contain the capacity factors for the scenarios "National Estimates 2025" and "National Estimates 2030"<br> - <em>PECD-2021.3-wide-Onshore-2025</em> and<em> PECD-2021.3-wide-Onshore-2030</em>: same as above but for wind onshore<br> - <em>PECD-2021.3-wide-Offshore-2025</em> and <em>PECD-2021.3-wide-Offshore-2030</em>: same as above but for wind offshore<br> - <em>PECD-wide-demand_national_estimates-2025</em> and<em> PECD-wide-demand_national_estimates-2030</em>: hourly electricity demand for all the climatic years for a specific zone. The two files contain the load for the scenarios "National Estimates 2025" and "National Estimates 2030" <br> - <em>PECD-2021.3-country-LFSolarPV-2025</em> and <em>PECD-2021.3-country-LFSolarPV-2030</em>: tables in Parquet format storing in each row the capacity factor for country/climatic year and hour of the year. The two files contain the capacity factors for the scenarios "National Estimates 2025" and "National Estimates 2030"<br> -<em> PECD-2021.3-country-Onshore-2025</em> and<em> PECD-2021.3-country-Onshore-2030</em>: same as above but for wind onshore<br> -<em> PECD-2021.3-country-Offshore-2025</em> and <em>PECD-2021.3-country-Offshore-2030</em>: same as above but for wind offshore<br> - <em>PECD-country-demand_national_estimates-2025</em> and <em>PECD-country-demand_national_estimates-2030</em>: same as above but for electricity demand<br> - <em>PECD_EERA2021_reservoir_pumping.zip</em>: archive with four files per each scenario: 1. table.csv with generation and storage capacities per zone/technology, 2. zone weekly inflow (GWh), 3. table.csv with generation and storage per country/technology and 4. country weekly inflow (GWh)<br> - <em>PECD_EERA2021_ROR.zip</em>: as for the previous file but the inflow is daily<br> - <em>plots.zip</em>: archive with 182 png figures with the weekly climatology for all the variables (daily for the electricity demand)</p> <p><strong>Note</strong></p> <p>I would like to thank Laurens Stoop for sharing the onshore wind data for the scenario 2030, that was corrupted in the original archive.</p>
Snapshot data in Parquet format from ENTSO-E Transparency Platform
<p>This is a snapshot of the following three datasets downloaded from the <a href="https://transparency.entsoe.eu/content/static_content/Static%20content/knowledge%20base/SFTP-Transparency_Docs.html">ENTSO-E Transparency Platform SFTP</a>:</p> <ol> <li>ActualTotalLoad (6.1.A)</li> <li>AggregatedGenerationPerType (16.1.B C)</li> <li>DayAheadPrices (12.1.D)</li> </ol> <p>The data has been downloaded and converted from CSV to Parquet using the package {arrow} in R.</p> <p><strong>WARNING</strong>: as <a href="https://github.com/energy-modelling-toolkit/entsoe-tp-survival-kit">I have explained here</a>, the data in the Transparency Platform changes every day. This means that the data contained in these files may be different from the data currently shown on the website or available via the APIs.</p>
Data from the ENTSO-E Transparency Platform - SQLite files
<p>Data extracted from the ENTSO-E Transparency Platform (transparency.entsoe.eu/).</p> <p> </p> <p>## AggregatedGenerationPerType (UPDATED 5/1/2021)</p> <p>Data in SQLite format. Timestamps (fields Datetime and Submission) are stored as seconds since 1970-1-1.</p> <p>CREATE TABLE gen_data<br> (Datetime INTEGER,<br> Submission INTEGER,<br> Resolution TEXT,<br> Area TEXT,<br> Type TEXT,<br> Generation REAL,<br> Consumption REAL,<br> PRIMARY KEY (Datetime, Area, Type))<br> </p>
GridKit extract of ENTSO-E interactive map
<p>This dataset was generated based on a map extract from May 11, 2016. This is an <em>unofficial</em> extract of the ENTSO-E interactive map of the European power system (including to a limited extent North Africa and the Middle East). The dataset has been processed by GridKit to form complete topological connections. This dataset is neither approved nor endorsed by ENTSO-E.</p> <p>This dataset may be inaccurate in several ways, notably:</p> <ul> <li>Geographical coordinates are transfered from the ENTSO-E map, which is known to choose topological clarity over geographical accuracy. Hence coordinates will not correspond exactly to reality.</li> <li>Voltage levels are typically provided as ranges by ENTSO-E, of which the lower bound has been reported in this dataset. Not all lines - especially DC lines - contain voltage information.</li> <li>Line structure conflicts are resolved by picking the first structure in the set</li> <li>Transformers are <em>not present</em> in the original ENTSO-E dataset, there presence has been derived from the different voltages from connected lines.</li> <li>The connection between generators and busses is derived as the geographically nearest station at the lowest voltage level. This information is again not present in the ENTSO-E dataset.</li> </ul> <p>All users are advised to exercise caution in the use of this dataset. No liability is taken for inaccuracies.</p>
ENTSO-E ERAA 2023 files in Parquet format
<p>demand time-series, PECD (renewables) and hydropower data from the ERAA 2023 (<a href="https://www.entsoe.eu/outlooks/eraa/2023/eraa-downloads/">ERAA Downloads | ENTSO-E – ERAA 2023 (entsoe.eu)</a>). </p> <p>Original data (in CSV and Excel format) is converted in Parquet in a format more suitable for modelling and analysis.</p> <p>This dataset includes:</p> <ul> <li>Demand time-series for all the modelled zones for the years 2025, 2028, 2030, 2033</li> <li>Capacity factors for all the modelled zones for the following technologies: wind onshore, wind offshore, solar PV rooftop, solar PV utility-scale, CSP with storage and without storage. All the capacity factors are available for the year 2025, 2028, 2030, 2033. All the files are available in the zipped file PECD.zip</li> <li>Hydropower data (installed capacity and time-series of daily/weekly inflow, min/max generating power, min/max reservoir levels) for run-of-river, pondage, reservoir-based plants, pumped-storage open- and closed-loop.</li> </ul>
ENTSO-E Electricity Demand - 01/2016 - 09/2022
<p>Electricity demand for ten European countries downloaded from the ENTSO-E Transparency Platform the 1st October 2022. The data covers the period 01/01/2016 - 31/09/2022. Saved in Parquet format.</p>
ENTSO-E PECD (European Climate Database) from MAF 2019 in CSV and Feather formats
<p>ENTSO-E has published the PECD dataset with the Mid-term Adequacy Forecast (MAF) 2019: https://www.entsoe.eu/outlooks/midterm/#download</p> <p>The downloadable archive contains 3 large Microsoft Excel (~350 MB each). Each file contains the hourly data for all the considered weather years (1982-2016) for all the MAF regions (grouped in tabs).</p> <p>Here we provide the same data but in a more open and user-friendly format.</p> <p><strong>Single files</strong></p> <p>Data saved in CSV and <a href="https://blog.rstudio.com/2016/03/29/feather/">Feather format</a> (using R package feather 0.35). Each file contains the hourly data in a wide tabular format with area, day, month, hour and year (1982-2016). The files are:</p> <ul> <li>PECD-MAF2019-wide-PV.csv (and .feather)</li> <li>PECD-MAF2019-wide-WindOffshore.csv (and .feather)</li> <li>PECD-MAF2019-wide-WindOnshore.csv (and .feather)</li> </ul> <p><strong>Split by year</strong></p> <p>We provide here three archives each one containing a file per year. The files are:</p> <ul> <li>PECD-MAF2019-wide-PV.single_years.tar.bz2</li> <li>PECD-MAF2019-wide-WindOffShore.single_years.tar.bz2</li> <li>PECD-MAF2019-wide-WindOnshore.single_years.tar.bz2</li> </ul> <p> </p> <p> </p>
ENTSO-E Hydropower modelling data (PECD) in CSV format
<p>PECD Hydro modelling</p> <p>This repository contains a more user-friendly version of the <code>Hydro modelling data</code> released by ENTSO-E with their <a href="https://www.entsoe.eu/outlooks/seasonal/">latest Seasonal Outlook</a>.</p> <p>The original URLs:</p> <ul> <li>The zipped file: <a href="https://eepublicdownloads.blob.core.windows.net/public-cdn-container/clean-documents/sdc-documents/seasonal/SOR2020/data/Hydro.zip">https://eepublicdownloads.blob.core.windows.net/public-cdn-container/clean-documents/sdc-documents/seasonal/SOR2020/data/Hydro.zip</a></li> <li>The documentation file (v 1.0): <a href="https://eepublicdownloads.blob.core.windows.net/public-cdn-container/clean-documents/sdc-documents/MAF/2019/Hydropower_Modelling_New_database_and_methodology.pdf">https://eepublicdownloads.blob.core.windows.net/public-cdn-container/clean-documents/sdc-documents/MAF/2019/Hydropower_Modelling_New_database_and_methodology.pdf</a></li> </ul> <p>The original ENTSO-E hydropower dataset integrates the PECD (Pan-European Climate Database) released for the <a href="https://www.entsoe.eu/outlooks/midterm/#download">MAF 2019</a></p> <p>As I did for the <a href="https://zenodo.org/record/3702418">wind & solar data</a>, the datasets released in this repository are <strong>only</strong> a more user- and machine-readable version of the original Excel files. As avid user of ENTSO-E data, with this repository I want to share my data wrangling efforts to make this dataset more accessible.</p> <p><strong>Data description</strong></p> <p>The <a href="https://eepublicdownloads.blob.core.windows.net/public-cdn-container/clean-documents/sdc-documents/seasonal/SOR2020/data/Hydro.zip">zipped file</a> contains 86 Excel files, two different files for each ENTSO-E zone.</p> <p>In this repository you can find 6 CSV files:</p> <ul> <li><code>PECD-hydro-capacities.csv</code>: installed capacities</li> <li><code>PECD-hydro-weekly-inflows.csv</code>: weekly inflows for reservoir and open-loop pumping</li> <li><code>PECD-hydro-daily-ror-generation.csv</code>: daily run-of-river generation</li> <li><code>PECD-hydro-weekly-reservoir-min-max-generation.csv</code>: minimum and maximum weekly reservoir generation</li> <li><code>PECD-hydro-weekly-reservoir-levels.csv</code>: weekly reservoir levels</li> <li><code>PECD-hydro-weekly-reservoir-min-max-uniform-levels.csv</code>: weekly minimum and maximum reservoir levels to use outside the climate years</li> </ul> <p><strong>Capacities</strong></p> <p>The file <code>PECD-hydro-capacities.csv</code> contains: run of river capacity (MW) and storage capacity (GWh), reservoir plants capacity (MW) and storage capacity (GWh), closed-loop pumping/turbining (MW) and storage capacity and open-loop pumping/turbining (MW) and storage capacity. The data is extracted from the Excel files with the name starting with <code>PEMM</code> from the following sections:</p> <ul> <li>sheet <code>Run-of-River and pondage</code>, rows from 5 to 7, columns from 2 to 5</li> <li>sheet <code>Reservoir</code>, rows from 5 to 7, columns from 1 to 3</li> <li>sheet <code>Pump storage - Open Loop</code>, rows from 5 to 7, columns from 1 to 3</li> <li>sheet <code>Pump storage - Closed Loop</code>, rows from 5 to 7, columns from 1 to 3</li> </ul> <p><strong>Inflows</strong></p> <p>The file <code>PECD-hydro-weekly-inflows.csv</code> contains the weekly inflow (GWh) for the climatic years 1982-2017 for reservoir plants and open-loop pumping. The data is extracted from the Excel files with the name starting with <code>PEMM</code> from the following sections:</p> <ul> <li>sheet <code>Reservoir</code>, rows from 13 to 66, columns from 16 to 51</li> <li>sheet <code>Pump storage - Open Loop</code>, rows from 13 to 66, columns from 16 to 51</li> </ul> <p><strong>Daily run-of-river</strong></p> <p>The file <code>PECD-hydro-daily-ror-generation.csv</code> contains the daily run-of-river generation (GWh). The data is extracted from the Excel files with the name starting with <code>PEMM</code> from the following sections:</p> <ul> <li>sheet <code>Run-of-River and pondage</code>, rows from 13 to 378, columns from 15 to 51</li> </ul> <p><strong>Miminum and maximum reservoir generation</strong></p> <p>The file <code>PECD-hydro-weekly-reservoir-min-max-generation.csv</code> contains the minimum and maximum generation (MW, weekly) for reservoir-based plants for the climatic years 1982-2017. The data is extracted from the Excel files with the name starting with <code>PEMM</code> from the following sections:</p> <ul> <li>sheet <code>Reservoir</code>, rows from 13 to 66, columns from 196 to 231</li> <li>sheet <code>Reservoir</code>, rows from 13 to 66, columns from 232 to 267</li> </ul> <p><strong>Reservoir levels</strong></p> <p>The file <code>PECD-hydro-weekly-reservoir-levels.csv</code> contains the minimum, maximum and the exact reservoir levels at beginning of each week (scaled coefficient from 0 to 1) for each climate year. The data is extracted from the Excel files with the name starting with <code>PEMM</code> from the following sections:</p> <ul> <li>sheet <code>Reservoir</code>, rows from 13 to 66, column 340 to 375</li> <li>sheet <code>Reservoir</code>, rows from 13 to 66, column 376 to 411</li> <li>sheet <code>Reservoir</code>, rows from 13 to 66, column 412 to 447</li> </ul> <p><strong>Reservoir levels</strong></p> <p>The file <code>PECD-hydro-weekly-reservoir-min-max-uniform-levels.csv</code> contains the minimum, maximum and the exact reservoir levels at beginning of each week (scaled coefficient from 0 to 1). The number are supposed to be used when climate years cannot be used (e.g. outside the range 1982-2017). The data is extracted from the Excel files with the name starting with <code>PEMM</code> from the following sections:</p> <ul> <li>sheet <code>Reservoir</code>, rows from 14 to 66, column 12</li> <li>sheet <code>Reservoir</code>, rows from 14 to 66, column 13</li> </ul> <p><strong>CHANGELOG</strong></p> <p>[2020/08/14] Added missing inflows for some countries (including Norway)<br> [2020/07/20] The old reservoir levels have been renamed 'uniform' consisting with the PECD source data. Added min, max and exact levels<br> [2020/07/17] Added maximum generation for the reservoir</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.