Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

6,766

datasets available to search

ShareScore release 0.7.1

Reset

Dataset results

6,766 results for “project”

Learn how ShareScore rates datasets ↗
zenodo48/100

Carbon Price Scenarios: Projecting prices for emission certificates

<p>This dataset consists of three different carbon price development scenarios. Each is represented by two growth rates which results in a total of 6 time series. The time frame is from 2020 to 2050. The units of the values are given in &euro; / t CO₂. All values are nominal.</p> <p>Overall, it should be noted that an estimate of the development of CO2 prices in the german nEHS and EU-ETS is subject to great uncertainty due to the major influence of regulatory intervention, a less liquid market towards 2030 and a lack of markets after 2030.</p> <p>The data provided is delivered in frictionless data format (see 2024-03-25_metadata_carbon-price-scenarios.package.json) and can be accessed using the frictionless software (https://frictionlessdata.io/).</p>

opencc-by-4.0Mar 2024View details →
zenodo48/100

Hydroelastic response of the scaled model of a floating offshore wind turbine platform in waves: HELOFOW Project Database

<p>This dataset contains the data measured during the <strong>HELOFOW </strong>model test campaign, performed at the Ocean and Hydrodynamic Engineering wave tank of Ecole Centrale Nantes (ECN): decay tests, regular wave tests and irregular waves tests. The preprocessed measured data is contained in MAT files.</p> <p>The model, the measurements and the tests are described in the appended Excel files.&nbsp;A Matlab(R) function is given as a short example to show how the MAT files are structured and how data may be handled for a plot.&nbsp;</p> <p>As stated in the reference paper (Leroy et al., <em>Ocean Engineering</em>, 2022):</p> <p>"As the size of floating wind turbines continues to increase, floating platforms reach dimensions that make their elastic and hydro-elastic behaviour significant. Several works in connection with the numerical modelling of the elastic behaviour of these wind turbines have been carried out but few validation data are available. This study focuses on the hydro-elastic response of a large floating wind turbine, in regular waves and severe sea-states. A new experimental wind turbine model has been designed to represent a 1:40 Froude-scaled spar platform carrying the DTU 10 MW turbine. The main challenge is here to reproduce a 1st bending mode frequency and hydrodynamic loads representative of a realistic large floating wind turbine. The platform model is made of a flexible backbone, reproducing the correct flexibility, and light floaters fixed on it provide the correctly scaled geometry. This experimental model is tested in various conditions including regular waves of several periods and steepness, and irregular waves of various intensity, including extreme 50-year return period conditions."</p> <p>&nbsp;</p> <p>This work was carried out within the framework of the WEAMEC, West Atlantic Marine Energy Community, and with funding from the Pays de la Loire Region and Europe (European Regional Development Fund).&nbsp;<br><br>HELOFOW project on <a href="https://www.weamec.fr/en/projects/helofow/">the WEAMEC website</a>.&nbsp;</p>

opencc-by-4.0Feb 2022View details →
zenodo48/100

Wave basin tests of multi-body floating photovoltaics system and an external floating breakwater.(SUREWAVE project)

<p><span>The aim of the EU Horizon Europe project SUREWAVE (2022-2025) is to develop a floating PV solution for offshore environments. A concrete floating breakwater (FBW) configuration will be designed to provide shelter for the floating PV (FPV) against harsh environmental conditions. MARIN&rsquo;s scope is to support the hydrodynamic design of the system through numerical simulations and wave basin tests. Basin tests are scheduled at two stages of the project: (1) at early design stage (for a global understanding of the preliminary design); (2) at final design stage (for verification and demonstration). The present dataset contains the reuslts of the early stage design stage wave basin testing.</span></p>

opencc-by-4.0Apr 2024View details →
zenodo48/100

Crop and soil measurements of quinoa in Morocco and Belgium (SALAD project)

<p>Crop and soil measurements of quinoa used to calibrate the SWAP-WOFOST model for the SALAD project (https://www.saline-agriculture.com/en).</p> <p>The data were collected from two locations:</p> <p>1) Laayoune, Southern Morocco: ICBA-Q5 quinoa variety, grown in 2021 under irrigation with saline water at levels of 4, 12, and 20 dS/m (https://doi.org/10.3389/fpls.2023.1143170)</p> <p><br>2) Merelbeke, Belgium: Bastille quinoa variety, grown in 2018, 2019, 2022, and 2023 under rainfed and non-saline conditions (https://www.quinoalokaal.be/nl/, https://doi.org/10.3390/plants10122689, https://doi.org/10.3390/plants11030265)</p> <p>&nbsp;</p>

opencc-by-4.0Nov 2024View details →
zenodo48/100

Topographic data and orthomosaics of the Noordwest Natuurkern project

<p>This data set contains digital elevation models and orthomosaics of the Noordwest Natuurkern region near Bloemendaal aan Zee, the Netherlands. High-frequency wind and rain data are also provided. The data descriptor paper can be found <a href="https://doi.org/10.3390/data9020037" target="_blank" rel="noopener">here</a>, and provides full details of data collection and processing, file format, file naming, and so on. The point clouds from which the elevation models were computed, are not available in this data base but can be obtained from the data creator upon reasonable request.</p> <p>v2.1.0: 41 digital elevation models (20080510 - 20241112; 1x1 m) and 27 orthomosaics (20130501 - 20241112; 1x1 and 0.05x0.05 m) - meta data file - wind and rain data - updated from v2.0.0 on January 8, 2025 (see ChangeLog.txt)</p> <p>&nbsp;</p>

opencc-by-4.0Aug 2022View details →
zenodo48/100

Data from the project "Nationalutgåva av de äldre geometriska kartorna", 1630–1655

<p>The project for the oldest geometrical maps, which took place between 2001 and 2010, aimed to make the content of these maps accessible to both the scientific community and the public.</p> <p>The large-scale maps were created during the period 1630&ndash;1655. The largest collection is held at the Swedish National Archives, within the Land Survey's archives. Additional maps are scattered across other public archives and private ownerships. The total number of map images is approximately 12,000.</p> <p>The project involved a comprehensive inventory of the archives, leading to a significant number of new discoveries. The original maps were scanned by the National Archives' photo unit. Each map image was printed as a scaled copy on archival paper. Following this, all text on the maps was interpreted, transcribed, systematized, and registered within a dedicated database. Additionally, the content on the maps regarding farms and settlements were located and spatially registered as point objects within a GIS application.</p> <p>The "National Edition of the Oldest Geometrical Maps" project was undertaken on behalf of Vitterhetsakademiens kommitt&eacute; f&ouml;r historiska kartor. Funding was provided by the Royal Swedish Academy of Letters, History and Antiquities and Riksbankens Jubileumsfond (principal investigator: Clas Tollin). Additional funding was provided by the Swedish Research Council to process and publish the database online. The project collaborated with the Institute for Language and Folklore, the Swedish Land Survey (Lantm&auml;teriet), and the Swedish National Archives.</p>

opencc-by-4.0Jun 2024View details →
zenodo48/100

Survey Study about Motivation for Participants in Citizen Science Projects

<p>The survey study about motivation for participants in Citizen Science projects was designed within the ongoing H2020 project named <a href="https://actionproject.eu/">ACTION</a> (pArticipatory sCience Toolkit agaInst pollutiON). Volunteers participate to citizen science initiatives for multiple reasons: personal enjoyment, desire for improvement or achievement, establishment of personal relationships, care for the environment, etc.<br> Studying motivation and investigating the factors influencing people participation to citizen science projects is an essential aspect in the analysis of citizen science communities. Understanding the reasons that foster people to engage can support the successful design and implementation of effective participant involvement tasks, as well as pave the way for long-term engagement.<br> The goal of the study is to analyse the motivation to participate of a specific citizen science community and the structure of the survey proposed shuold be customised considering the topic of the activities.</p> <p><br> This research object describes the studies performed within the ACTION project to investigate motivations of different citizen science communities: the TESS Network (<a href="https://tess.stars4all.eu/">https://tess.stars4all.eu/</a>); the 6 ACTION pilots (<a href="https://actionproject.eu/citizen-science-pilots">https://actionproject.eu/citizen-science-pilots</a>) Mapping Mobility, Open Soil Atlas, Water Sentinels, Restart Data Workbench, Wow Nature, Walk Up Aniene.<br> The surveys were designed and administered through Coney (<a href="https://coney.cefriel.com">https://coney.cefriel.com</a>) and made available as linked data exploiting the Survey Ontology (<a href="https://w3id.org/survey-ontology">https://w3id.org/survey-ontology</a>).</p> <p>The research object adopts the <a href="https://www.researchobject.org/ro-crate/1.0/">RO-Crate</a>&nbsp;specification. Files made available within the research object&nbsp;are:</p> <ul> <li><em>*-procedure.ttl</em>&nbsp;contains the RDF representation of the <strong>template</strong> structure of the&nbsp;conversational survey (questions, answers, etc.)&nbsp;using the <a href="https://w3id.org/survey-ontology">Survey Ontology</a></li> <li>*-<em>mean-var-motivating-questions.csv </em>contains&nbsp;the computed mean and average for each question considered (observable variables) comparing all the surveys performed</li> <li>*-<em>mean-var-motivating-factor.csv </em>contains the computed mean and average for each motivation factor considered (latent variables) comparing all the surveys performed</li> <li>*-<em>correlation-factors-global-motivation.csv </em>contains the correlation analysis between each motivation factor and the global motivation&nbsp;comparing all the surveys performed</li> </ul> <p>The Research Object also references all the RO-Crates describing the different survey motivation studies in details.</p>

opencc-by-4.0Dec 2021View details →
zenodo48/100

A dataset of 150000 terminal weighted projective spaces

<p><strong>Weighted projective spaces with at worst terminal singularities</strong></p> <p>A dataset of 150000 randomly generated weighted projective spaces with at worst terminal singularities, in dimensions 1 to 10.</p> <p>The data consists of the plain text files &quot;rank_1_dim_N.txt&quot; where N, which is the dimension of the weighted projective space, is in the range 1 to 10. Each line of the file is a sequence of weights of length N+1. For example, the first line of &quot;rank_1_dim_4.txt&quot; is:</p> <p>[1,2,5,14,21]</p> <p>and this corresponds to the 4-dimensional weighted projective space P(1,2,5,14,21).</p> <p>For details, see the paper:</p> <p>&quot;Machine learning the dimension of a Fano variety&quot;, Tom Coates, Alexander M. Kasprzyk, and Sara Veneziale,&nbsp;<em>Nature Communications</em>, <strong>14:</strong>5526&nbsp;(2023). doi:10.1038/s41467-023-41157-1</p> <p>Magma code capable of generating this dataset is in the file &quot;generate_rank_1.m&quot;.</p> <p>If you make use of this data, please cite the above paper and the DOI for this data:</p> <p>doi:10.5281/zenodo.5790079</p>

opencc-zeroJan 2022View details →
zenodo48/100

GERONTE H2020 project - GERDAT005 - Intrinsic capacity evaluation and intervention protocol

<p><strong>The present document is a dataset generated as part of Deliverable D1.1.&nbsp;of the GERONTE project, which has received funding from the European Union&rsquo;s Horizon 2020 Programme under Grant Agreement N&deg;945218. It aims to provide the geriatric oncology professional community with a&nbsp;protocol for the evaluation of intrinsic capacity and frailty, with subsequent interventions aimed at optimizing health status and support for older patients with multimorbidity and cancer.</strong></p> <p>GERONTE is a 5-year research and innovation project (April 2021 to Mars 2026) funded by the European Union within the framework of the H2020 Research and Innovation programme, in response to the health societal challenge topic SC1-BHC-24-2020 &ldquo;Healthcare interventions for the management of the elderly multimorbid patient&rdquo;. The overall aim of GERONTE is to improve quality of life - defined as well-being on three levels: global health status, physical functioning and social functioning- for older multimorbid patients, while reducing overall costs of care. To this end, GERONTE will co-design, test, and prepare for deployment an innovative cost-effective patient-centred holistic health management system, hereafter referred to as the GERONTE intervention. GERONTE intervention will rely on an ICT based application for real-time collection and integration of standardised clinical and home patient-reported data. GERONTE intervention will be demonstrated in the context of care of multimorbid patients having cancer as a dominant morbidity, and be adaptable to any other combination of morbidities.</p> <p>An important component of Geronte is to take account of intrinsic capacity. Most older patients who are diagnosed with cancer also suffer from other illnesses and impairments that could affect their prognosis, priorities and ability to tolerate and benefit from treatment. For tailored oncologic decision making, it is essential to obtain a complete overview of the patient&rsquo;s health status. This dataset contains the protocol for evaluation intrinsic capacity and potential interventions for impairments or vulnerabilities that were identified in this evaluation. It is going to be used in the assessment and management of older patients with multimorbidity and cancer, within the GERONTE care pathway.</p>

opencc-by-4.0Mar 2022View details →
zenodo48/100

In situ smartphone radiometry of Lake Balaton, MONOCLE H2020 project

<p>This dataset includes in situ radiometric data collected from Lake Balaton and Kis-Balaton, Hungary&nbsp;during July 3<sup>rd</sup>&nbsp;&ndash; July 5<sup>th</sup>, 2019 within the framework of Horizon 2020 MONOCLE project.</p> <p>The Python code used to analyse the data and produce the summary tables, and which should be used to read the data in, can be found here:&nbsp;<a href="https://github.com/burggraaff/smartphone-water-colour">https://github.com/burggraaff/smartphone-water-colour</a></p> <p>The data are structured as follows:</p> <p>&quot;Balaton_2019070x&quot; - These folders contain the RAW and JPEG smartphone images, sorted by station, time, and smartphone. Each low-level subfolder includes the RAW and JPEG images and a CSV file with the derived radiance, R_rs, and other values.</p> <p>&quot;Discarded_data&quot; - These folders contain RAW and JPEG smartphone images that were not used, organised in the same way as &quot;Balaton_2019070x&quot;. They are included here for completeness. Each subfolder contains a file explaining why the data were not used.</p> <p>&quot;Greycard&quot; - This folder contains the RAW smartphone images used to characterise the angular response of the grey card. Also included are the derived mean/uncertainty values in .NPY format.</p> <p>&quot;BALATON_2019_STATION_LOG&quot; - This worksheet contains the station log of the MONOCLE field campaign at Lake Balaton in 2019, during which our field data were taken.</p> <p>&quot;balaton_xxx_18pct.csv&quot; - These CSV files contain summaries of the data derived from the smartphone RAW images, including radiance, R_rs, etc. They are most easily read in using the Python code linked above. These files were generated by stacking the individual CSV files from each station/time/smartphone. The filename indicates the smartphone and data type.</p> <p>&quot;balaton_Samsung_Galaxy_S8_raw_replicates.csv&quot; - This CSV file contains the relative uncertainty (R/G/B and band ratios, in %) and absolute uncertainty (hue angle and FU) in replicate Galaxy S8 measurements, as described in the paper.</p> <p>&quot;greycard_data_Maine.csv&quot; - This CSV file contains the data used to determine the spectral response of the grey card. It includes spectral measurements of the surface irradiance on a white panel and the grey card, and the downwelling irradiance measured with a cosine collector.</p> <p>&quot;README.txt&quot; - Description of the data.</p> <p>&quot;So-Rad_Balaton2019.csv&quot; - This CSV file contains the (ir)radiance data and R_rs from the So-Rad on 3 July 2019, processed using the 3C method, as described in the paper.</p> <p>&quot;wisp_Balaton_20190703_20190705_table.csv&quot; - This CSV file contains the (ir)radiance data and R_rs from the WISP-3 on 3-5 July 2019, processed using the Mobley method, as described in the paper.<br> &nbsp;</p>

opencc-by-4.0May 2022View details →
zenodo48/100

New maps of global geologic provinces and tectonic plates: global tectonics data and QGIS project file

<p>The global tectonics data compilation is a set of raster and vector data that are useful for investigating tectonics past and present. &nbsp;The datasets are useful on their own or can be used in GIS software, which includes the QGIS project file for convenience. &nbsp;The datasets include our new models for tectonic plate boundaries and deformation zones, geologic provinces and orogens. &nbsp;Additional datasets include earthquake and volcano locations, geochronology, topography, magnetics, gravity, and seismic velocity.</p> <p>The global tectonics collection is suitable for research and educational purposes.</p>

opencc-by-4.0Mar 2022View details →
zenodo48/100

Field data obtained from the implementation of the Field Protocols (1, 2, and 3) from WP3 in B-GOOD Project

<p>Dataset of the field data obtained from the implementation of the field protocols 1, 2, and 3, developed under B-GOOD (Giving Beekeeping Guidance by cOmputatiOnal-assisted Decision making, Grant agreement No. 81762) project&nbsp;in WP3.</p> <p>The primary goal of field data collection was to gather data that could be used for the validation of floral resources maps and models developed across the various tasks within WP3, and also to fill specific information gaps during the development of the phenological model. The fieldwork was conducted in three countries (Portugal, Belgium, and the United Kingdom) using three field protocols developed under B-GOOD and described in Milestone MS15.</p> <p>These protocols are part of Tasks 3.1 and 3.3 and serve various purposes. The primary goal of Field Protocol 1: &quot;<em>Assessment of plant species composition on key landscape elements/habitats important for bees</em>&quot; was to determine the species composition in selected key plant communities (<em>i.e.</em> landscape elements or habitats) important for bees. Field Protocol 1 was used to determine the plant species composition of specific ALMaSS landscape elements and to confirm/validate the plant composition of some BIOEUNIS habitat types, developed in Task 3.1. The main goal of Field Protocol 2: &quot;<em>Assessment of Phenology of Floral Resources for Bees</em>&quot; was to determine the flower phenology of targeted plant species to construct flowering phenological curves for targeted plant species. Field Protocol 2 was used to validate the floral resource models developed in Task 3.2 (see Section 2.4). The primary goal of Field Protocol 3: &quot;<em>Floral resources evaluation (detailed method)</em>&quot; was to conduct a detailed evaluation of the floral resources in each landscape window with B-GOOD mini-apiaries to map resource availability. Field protocol 3 was divided into two parts: Part 1 - &ldquo;<em>Assessment and quantification of floral resources</em>&rdquo;, aiming to determine the species composition, species cover, and flower abundance; and Part 2 - &ldquo;<em>Flowering species characterization</em>&rdquo;, aiming to quantify the number of flowers per individual plant, and the nectar and pollen production of target plant species. Field Protocol 3 was also used to make a detailed assessment of plant species composition and plant resources at the landscape level, as well as to fill the gaps in knowledge about pollen and nectar production of some target plant species. The field data collected by the implementations of the field protocols could be categorized into three main groups: Plant Species Composition (Field Protocol 1 and Field Protocol 3: Part 1), Phenology of Floral Resources (Field Protocol 2), and Flowering Species Characterization (Field Protocol 3: Part 2).</p> <p>Field protocols have been implemented in Portugal, the United Kingdom, and Belgium. Field protocols 2 and 3 were fully implemented in the three countries. However, field protocol 1 was not implemented as a stand-alone field protocol in Portugal and the United Kingdom due to logistical and time constraints primarily caused by the COVID pandemic. However, this does not hamper our ability to obtain landscape-specific plant composition data because the information gathered in this protocol can be derived entirely from the implementation of the first part of Field Protocol 3. As a result, for Portugal and the United Kingdom, Field Protocol 1 data was derived from the first part of Field Protocol 3. The full dataset gathered by the implementation of these three protocols is available here.</p> <p>The files &ldquo;field-data-protocol-1-be.xlsx&rdquo;, &ldquo;field-data-protocol-1-pt.xlsx&rdquo; and &ldquo;field-data-protocol-1-uk.xlsx&rdquo; have the field data obtained from the implementation of Field Protocol 1: &quot;Assessment of plant species composition&quot; in Belgium, Portugal, and the UK, respectively.</p> <p>The files &ldquo;field-data-protocol-2-be.xlsx&rdquo;, &ldquo;field-data-protocol-2-pt.xlsx&rdquo; and &ldquo;field-data-protocol-2-uk.xlsx&rdquo; have the field data obtained from the implementation of Field Protocol 2: &quot;Assessment of Phenology of Floral Resources&quot; in Belgium, Portugal, and the UK, respectively.</p> <p>The files &ldquo;field-data-protocol-3-part-1-be.xlsx&rdquo;, &ldquo;field-data-protocol-3-part-1-pt.xlsx &ldquo;and &ldquo;field-data-protocol-3-part-1-uk.xlsx&rdquo; have the field data obtained from the implementation of Field Protocol 3 - Part 1 &ldquo;Assessment and quantification of floral resources&rdquo; in Belgium, Portugal, and the UK, respectively.</p> <p>The file &ldquo;field-data-protocol-3-part-2-be-pt-uk.xlsx&rdquo; have the field data obtained from the implementation of Field Protocol 3 - Part 2 - &ldquo;Flowering species characterization&rdquo; in Belgium, Portugal, and the UK.</p> <p>For further details, see Alves da Silva et al. 2020. Field protocols for the assessment of Floral Resources. Milestone MS15 EU Horizon 2020 B-GOOD Project. GA No. 817622 and&nbsp;Zi&oacute;łkowska et al 2022. Floral Resource Models Validation Deliverable D3.4 EU Horizon 2020 B-GOOD Project, GA No. 817622.</p>

opencc-by-4.0May 2022View details →
zenodo48/100

Sonnets base of the Oupoco project

<p>The Oupoco Database is a collection of 4870 French sonnets developed in the framework of the Oupoco Project. The database is mainly composed of poems from the 19th and early 20th century. We have identified 767 authors: 4412 sonnets written by men (660), 439 sonnets written by women (107), which leaves 19 sonnets for which we have not been able to assign a female or male author. The sonnets come from different sources from the Internet, or not: we especially want to thank the Biblioth&egrave;que nationale de France (the French national library) that gave us access to a large corpus, from which we were able to extract an invaluable number of great poems. To all the sonnets is attached a specific license related to the source they come from, but all are freely available and can be re-used for free. This database has initially been developed for the Oupoco project (L&#39;Ouvroir de litt&eacute;rature combinatoire, https://oupoco.org/), which consists in producing new sonnets by recombining verses from existing ones from the French literature, following the idea put forward by Queneau in his famous conceptual book: Cent mille milliards de po&egrave;mes (1961). Different scripts have been developed for the Oupoco project (to analyse the rhymes and recombine the verses) which are not part of this data base but can be obtained by contacting the authors of the project. Beyond Oupoco, this database can be used for various purposes, for teaching and for research, especially in the following domains: literature studies, corpus linguistics, digital humanities, arts and technology, etc.&nbsp;</p>

opencc-by-4.0Jun 2022View details →
zenodo48/100

Citizen Science projects on Alien Species in Europe

<p><strong>Context</strong></p> <p>This survey relates to COST (European Cooperation in Science and Technology) Action CA17122 - Alien CSI - Increasing understanding of alien species through citizen science (see https://alien-csi.eu/). The main aim of this survey was&nbsp;to collect information on Citizen Science projects/initiatives involving alien species in European Member States and some neighbouring countries. The survey was performed using a google forms. Survey respondents/contributors&nbsp;are mentioned in this dataset as data collectors.&nbsp;</p> <p><strong>Definitions</strong></p> <p>We defined Citizen Science projects as project which actively involved citizens in scientific enquiry generating new knowledge or understanding on alien species. Citizens may act as contributors, collaborators, or as project leader and have a meaningful role in the project.&nbsp;&#39;Alien Species&#39; are defined as&nbsp;any &nbsp;live &nbsp;specimen &nbsp;of &nbsp;a &nbsp;species, &nbsp;subspecies &nbsp;or &nbsp;lower &nbsp;taxon &nbsp;of &nbsp;animals, &nbsp;plants, &nbsp;fungi or &nbsp;micro-organisms &nbsp;introduced &nbsp;outside &nbsp;its &nbsp;natural &nbsp;range; &nbsp;it &nbsp;includes &nbsp;any &nbsp;part, &nbsp;gametes, &nbsp;seeds, &nbsp;eggs &nbsp;or &nbsp;propagules &nbsp;of &nbsp;such species, &nbsp;as &nbsp;well &nbsp;as &nbsp;any &nbsp;hybrids, &nbsp;varieties &nbsp;or &nbsp;breeds &nbsp;that &nbsp;might &nbsp;survive &nbsp;and &nbsp;subsequently &nbsp;reproduce. Alien Species thus includes both species that are invasive and species that are alien but not invasive. An&nbsp;&#39;Invasive Alien Species&#39; is defined as an alien species whose introduction or spread has been found to threaten or adversely impact &nbsp;upon biodiversity and/or related ecosystem services.</p> <p><strong>Survey methodology</strong></p> <p>The survey was made available on Google Forms and disseminated online, collecting responses from June 27, 2019 to April 6, 2020. It was shared with all COST Action CA17122 participants and in each country one person coordinated contacts with existing citizen science projects involving alien and/or invasive species and requested that they complete the survey. Thus, all projects were active in EU member states and neighbouring countries, though some may also be active outside of Europe. To increase reach, the survey was also disseminated through the European Citizen Science Association (ECSA) newsletter and mailing list and respondents were asked to share it with colleagues and local networks via snowball sampling.</p> <p><strong>Questions and attribute values</strong></p> <p>Survey questions and attribute values were developed using JRC metadata standards for CS projects (Bio Innovation Service 2018) and the project metadata model of PPSR Core, a set of global, transdisciplinary data and metadata standards for Public Participation in Scientific Research (https://core.citizenscience.org/). The survey included 62 questions in nine sections:</p> <ol> <li>Contact information of the respondent;</li> <li>General characterization of the project, including a brief summary, geographical scope, time scale, hosting entities, funding, etc.;&nbsp;</li> <li>Information on project scope, including target audience, taxonomic and environmental scope, project aims, type of data collected, etc.;</li> <li>Policy-related information, namely if the project has policy relevance and inclusion of species listed in the EU IAS Regulation;</li> <li>Information on engagement, such as type of involvement of citizens in the design of the project, engagement methods and social media used, skills needed to participate and frequency of contributions;</li> <li>Information on feedback and support provided to participants by the project, e.g., if projects provide materials for species identification, guidelines, training activities, information on how data from the project are used, feedback mechanisms and support;&nbsp;</li> <li>Data quality and data management, namely validation mechanism for records, registration type, methods of recording, whether data are open and accessible to citizen scientists, data form used to store data, data standards and data licence used, whether a public data management plan was drafted for the project, and the vocabulary used with respect to biological invasions (origin, occurrence status, degree of establishment and pathway of introduction);</li> <li>Performance indicators of projects, namely, usage of apps, number of participants and number of records, whether learning is assessed, number and type of publications using data from the project;&nbsp;</li> <li>Notes and remarks.</li> </ol> <p><strong>Files</strong></p> <ul> <li><strong>raw_data.xlsx</strong>: includes the non-processed survey responses, supplemented with a project_ID. All GDPR sensitive data such as email addresses were&nbsp;omitted. Each row represents one project.</li> <li><strong>projects_excluded.csv:&nbsp;</strong>includes all projects that were omitted from the analysis and the specific criteria for this exclusion.&nbsp;</li> <li><strong>processed_data.csv</strong>:&nbsp;includes the cleaned, processed survey responses, used for analysis. The R code used for the analysis is available on <a href="https://github.com/alien-csi/inventory-analysis/blob/master/src/analysis.Rmd">this&nbsp;github repository</a>.</li> <li><strong>survey.pdf:&nbsp;</strong>a pdf extract from the original Google Forms, including all questions and their specifications.&nbsp;</li> <li><strong>analysis.Rmd</strong>: Rmarkdown script for statistical analysis. Also available&nbsp;on <a href="https://github.com/alien-csi/inventory-analysis/blob/master/src/analysis.Rmd">this&nbsp;github repository</a>.</li> </ul> <p>&nbsp;</p> <p>&nbsp;</p>

opencc-by-4.0Apr 2021View details →
zenodo48/100

Descriptions of SNSF-funded research projects

<p>This repository contains the data to replicate the analyses performed in Meier, D. S., Mata, R., &amp; Wulff, D. U. (2021). text2sdg: An R package to Monitor Sustainable Development Goals from Text. <em>arXiv preprint arXiv:2110.05856</em>.</p> <p>The <a href="../api/records/11060662/draft/files/GrantWithAbstracts.csv/content" target="_blank" rel="noopener noreferrer">GrantWithAbstracts.csv</a> data is originally provided by the Swiss National Science Foundation and can also be downloaded from their <a href="https://data.snf.ch/datasets">website</a>. This data contains information on research projects funded by the Swiss National Science Foundation between 1975 and 2022. Among other things, the data provides information on what the funded projects were about and how much funding was provided.</p> <p>The <a href="../api/records/11060662/draft/files/backtrans_table.RDS/content" target="_blank" rel="noopener noreferrer">backtrans_table.RDS</a> data contains 1,500 randomly selected projects from the&nbsp;<a href="../api/records/11060662/draft/files/GrantWithAbstracts.csv/content" target="_blank" rel="noopener noreferrer">GrantWithAbstracts.csv</a> data that were translated from English to German and then from German back to English.</p> <p>The <a href="../api/records/11060662/draft/files/benchmark_table_revision.rds/content" target="_blank" rel="noopener noreferrer">benchmark_table_revision.rds</a> data contains runtimes from benchmarking the text2sdg R package.&nbsp;</p>

opencc-by-4.0Apr 2024View details →
zenodo48/100

Slovenský Supermodel P&T1 (SSPT1) : Matej Bel University SKRIPTOR project datasets

<p><strong>SLO:</strong></p> <p>Dňa 17.05.2024 sme spustili vo webovej aplik&aacute;cii Transkribus tvorbu nov&eacute;ho agregovan&eacute;ho slovensk&eacute;ho supermodelu. Z&aacute;klad pre tvorbu supermodelu pre určit&eacute; slovensk&eacute; tlačen&eacute; historick&eacute; dokumenty a strojom p&iacute;san&eacute; dokumenty tvorili parci&aacute;lne modely rie&scaron;iteľov &uacute;loh v projekte&nbsp; <strong>Skriptor</strong> (<strong>Univerzita Mateja Bela v Banskej Bystrici a &Scaron;t&aacute;tna vedeck&aacute; knižnica v Banskej Bystrici</strong>), ako aj transkripcie, ktor&eacute; pripravili &scaron;tudenti <strong>Slezskej univerzity v Opave v r&aacute;mci &Scaron;tudentskej grantovej s&uacute;ťaže</strong> a vzdel&aacute;vac&iacute;ch aktiv&iacute;t.&nbsp;</p> <p><strong>Michaela Miku&scaron;kov&aacute; a Lucia Nižn&iacute;kov&aacute;</strong> v r&aacute;mci projektu <strong>Skriptor</strong> kompletne spracovali n&aacute;ročn&uacute; segment&aacute;ciu a manu&aacute;lnu transkripciu <strong>92 s.</strong> GT historickej tlačenej knihy J.A. Komensk&eacute;ho&nbsp;<strong>Orbis Pictus</strong> (vydanie z roku 1798). I&scaron;lo, z hľadiska transkripcie o mimoriadne komplikovan&uacute; &uacute;lohu, pretože kniha m&aacute; mnoho ilustr&aacute;ci&iacute;, je p&iacute;san&aacute; v 4 jazkoch (latinčina, maďarčina, nemčina, če&scaron;tina), navy&scaron;e vo forme tabuliek a p&iacute;smom antikva a &scaron;vabach.&nbsp;</p> <p><strong>Du&scaron;an Katu&scaron;č&aacute;k </strong>v r&aacute;mci projektu <strong>Skriptor, </strong>vzdel&aacute;vac&iacute;ch aktiv&iacute;t a &scaron;tudentskej grantovej s&uacute;ťaže<strong> </strong>SGS na Slezskej univerzite a vedenia diplomovej pr&aacute;ce v Opave spracoval cel&yacute; do kvality GT cel&yacute; rad historick&yacute;ch nov&iacute;n, časopisov a kn&iacute;h z 19. a začiatku 20 storočia (Moravsk&eacute; noviny (1849), Programov&eacute; bulletiny Slovenskej filharm&oacute;nie (1849-1970), Opavsk&yacute; Besedn&iacute;k (1863), Jitrenka (1840), I. Palugyay: Kde jest pravda (1854), lužickosrbsk&yacute; časopis Lužica (1909), &Scaron;labik&aacute;r (1872), J.M. Hurban: Cirkev Ewanjelicko-Luther&aacute;nska (1861), J.N. Bobula: J&aacute;no&scaron;&iacute;k (1862), D. Lichard: Obzor (1866) a i. Niektor&eacute; dokumenty s&uacute; už kompletne transkribovan&eacute; použit&iacute;m priv&aacute;tnych modelov (ca 1000 s.), av&scaron;ak do datasetu SSPT1 boli použit&eacute; len sety GT.</p> <p><strong>Kl&aacute;ra Kov&aacute;čov&aacute;-Pohlov&aacute;</strong> (Diplomov&aacute; pr&aacute;ca, 2024, FPF SU Opava) a&nbsp;<strong>Matej &Scaron;mida</strong> (UMB Bansk&aacute; Bystrica) spracovali strojopisn&eacute; dokumenty, pričom použili vzorky r&ocirc;znych fontov v slovenskom, českom, nemeckom jazyku (ca 150 s.)</p> <p><strong>Nikola Halfarov&aacute;, Terezie Gajdo&scaron;ov&aacute;, Lenka M&aacute;lkov&aacute;, Nikol Taufrov&aacute;, Nela Koci&aacute;nov&aacute;</strong> (4. roč, FPF SLU)v predmete prof. Du&scaron;ana Katu&scaron;č&aacute;ka Digitalizace II. pripravili ca 80 s. prepisov GT z r&ocirc;znych historick&yacute;ch tlač&iacute; z 18. a 19. storočia p&iacute;san&yacute;ch v če&scaron;tine (&scaron;vabach).&nbsp;</p> <p>Model m&aacute; označenie&nbsp;<strong>ID78289 SLOVAK Supermodel print&amp;typewriter (SSPT1)</strong> &nbsp;sme použili&nbsp; <strong>542 str&aacute;n v kvalite Ground Truth (GT 37897 riadkov a 200697 slov). 59 str&aacute;n na overenie nov&eacute;ho modelu (Validation</strong> set <strong>)</strong> . repozit&aacute;rov &Scaron;t&aacute;tnej vedeckej knižnice v Ostrave, Slovenskej n&aacute;rodnej knižnice v Martine, z repozit&aacute;ra Manuskriptorium, zo &Scaron;t&aacute;tneho arch&iacute;vu v Banskej Bystrici a z Knižnice Univerzity Mateja Bela v Banskej Bystrici.&nbsp;<br>Samotn&eacute; uk&aacute;žky považujeme pre použ&iacute;vanie ďal&scaron;ieho a zdokonaľovania modelu za veľmi d&ocirc;ležit&yacute;, ďal&scaron;&iacute; v&yacute;skumn&iacute;ci dostan&uacute; predstavu o podobnom alebo odli&scaron;nom p&iacute;sme vlastn&yacute;ch dokumentov, ktor&eacute; chc&uacute; transkribovať.&nbsp;</p> <p><strong>V modeli ID78289 SLOVAK Supermodel print&amp;typewriter (SSPT1) boli dosiahnut&eacute; hodnoty Train set: 1,00% a Validation set: 1,00%. Znamen&aacute; to teda &bdquo;presnosť&ldquo; automatickej transkripcie 99%.</strong></p> <p>Tvorba modelu SSM1 na servri Transkribus trvala 21 hod&iacute;n a 52 min&uacute;t. Proces tvorby bol nastaven&yacute; na 100 cyklov a skončen&yacute; po&nbsp; <strong>100 cykloch</strong> (epoch).&nbsp; <br>Model SSPT1 je prv&yacute;m pokusom na Slovensku av Česku o tvorbe agregovan&eacute;ho n&aacute;stroja, prostredn&iacute;ctvom ktor&eacute;ho by bolo možn&eacute; automaticky spr&iacute;stupniť určit&eacute; typy tlačen&yacute;ch a strojopisn&yacute;ch dokumentov, ktor&eacute; s&uacute; podobn&eacute; p&iacute;smam použit&yacute;m v jeho tvorbe.&nbsp; <br>V pr&iacute;padn&yacute;ch nemožnoch <strong>ID78289 SLOVAK Supermodel print&amp;typewriter (SSPT1)</strong> považujeme za definit&iacute;vny univerz&aacute;lny model transkripcie historick&yacute;ch tlač&iacute; a strojopisov slovenskej proveniencie v&scaron;etk&yacute;ch typov a obdob&iacute;. Varieta p&iacute;siem a &scaron;t&yacute;lov je rozmanit&aacute; a tvorba optim&aacute;lneho agregovan&eacute;ho modelu predstavuje &uacute;lohu pre ďal&scaron;&iacute;ch v&yacute;skumn&iacute;kov a entuziastov v nasleduj&uacute;cich rokoch.&nbsp; <br>Domnievame sa v&scaron;ak, že n&aacute;&scaron; prv&yacute; agregovan&yacute; model <strong>SSPT1</strong> m&ocirc;že byť potrebn&yacute; automatick&uacute; transkripciu ďal&scaron;&iacute;ch anal&oacute;gov&yacute;ch dokumentov.&nbsp;<br>V&yacute;skumn&yacute; t&yacute;m pl&aacute;nuje spr&iacute;stupniť datasety v r&aacute;mci udržateľnosti projektu v roku 2024-2028 prednostn&eacute; pre v&yacute;skumn&eacute; a vzdel&aacute;vacie &uacute;čely pre in&scaron;tit&uacute;cie a v&yacute;skumn&iacute;kov, ktor&iacute; bud&uacute; chcieť prispieť k modelu historick&yacute;ch a nov&yacute;ch dokumentov v z&aacute;padoslovansk&yacute;ch jazykoch, resp. jazykov slovenskej a bohemik&aacute;lnej proveniencie.&nbsp; <strong>Copyright: CC BY-NC-SA.</strong><br>Samozrejme, tak&aacute;to automatick&aacute; transkripcia neprinesie hneď uspokojiv&eacute; v&yacute;sledky. M&ocirc;že v&scaron;ak byť &bdquo;hrub&uacute;&ldquo; postupn&uacute; automatick&uacute; transkripciu ďal&scaron;&iacute;ch str&aacute;n, ich manu&aacute;lnu opravu do stavu GT a n&aacute;sledn&eacute; použitie v&auml;č&scaron;&iacute;ch datasetov GT na zdokonalenie nov&eacute;ho modelu na b&aacute;ze n&aacute;&scaron;ho <strong>SSPT1.</strong> Po vytvoren&iacute; ďal&scaron;&iacute;ch stoviek a tis&iacute;cov str&aacute;n GT bude možn&eacute; prist&uacute;piť k tvorbe ďal&scaron;&iacute;ch gener&aacute;ci&iacute; nov&yacute;ch modelov na z&aacute;klade <strong>SSPT1</strong> . V&yacute;voj by mohol pokračovať pre tlač a strojopisy modelmi nov&yacute;ch gener&aacute;ci&iacute; <strong>SSPT2</strong> , <strong>SSPT3</strong> ap.</p> <p>V&yacute;zvu pre v&yacute;skumn&iacute;kov predstavuje aj v&yacute;voj a tvorbu&nbsp;<strong>nov&eacute;ho agegovan&eacute;ho supermodelu, ktor&yacute; by zahrnul jednak rukopisy a jednak tlače a strojopisy. Tento slovensk&yacute; supermodel by mohol byť zdieľan&yacute; v r&aacute;mci komunity odborn&iacute;kov Transkribus a zahrnut&yacute; do niektor&eacute;ho veľk&eacute;ho supermodelu Transkribus Community ap.&nbsp;</strong></p> <p>&nbsp;</p> <p><strong>ENG:&nbsp;</strong></p> <p>On May 17, 2024, we launched the creation of a new aggregated Slovak supermodel in the Transkribus web application. The basis for the creation of a supermodel for certain Slovak printed historical documents and typewritten documents was the partial models of task solvers in the Skriptor project (Matej Bela University in Bansk&aacute; Bystrica and the State Science Library in Bansk&aacute; Bystrica), as well as transcriptions prepared by students of the University of Silesia in Opava in within the Student Grant Competition and educational activities.</p> <p>As part of the Skriptor project, Michaela Miku&scaron;kov&aacute; and Lucia Nižn&iacute;kov&aacute; completely processed the demanding segmentation and manual transcription of 92 s. GT of historical printed book J.A. Comenius' Orbis Pictus (1798 edition). From the point of view of transcription, it was an extremely complicated task, because the book has many illustrations, it is written in 4 languages (Latin, Hungarian, German, Czech), in addition in the form of tables and in antique and Swabian script.</p> <p>Du&scaron;an Katu&scaron;č&aacute;k, as part of the Skriptor project, educational activities and the SGS student grant competition at the University of Silesia, and the management of the diploma thesis in Opava, processed a whole series of historical newspapers, magazines and books from the 19th and early 20th centuries (Moravsk&eacute; noviny (1849), Program bulletins of the Slovak Philharmonic (1849-1970), Opavsk&yacute; Besedn&iacute;k (1863), Jitrenka (1840), I. Palugya: Kde jest pravda (1854), Lusatian Serbian magazine Lužica (1909), &Scaron;labik&aacute;r (1872), J.M. Hurban: Cirkev Ewanjelicko- Luther&aacute;nska (1861), J.N. Bobula: J&aacute;no&scaron;&iacute;k (1862), D. Lichard: Obzor (1866) and others. Some documents are already completely transcribed using private models (about 1000 pages), but only GT sets were used.</p> <p>Kl&aacute;ra Kov&aacute;čov&aacute;-Pohlov&aacute; (Diplomov&aacute; pr&aacute;ce, 2024, FPF SU Opava) and Matej &Scaron;mida (UMB Bansk&aacute; Bystrica) processed typewritten documents, using samples of various fonts in Slovak, Czech, and German languages (ca. 150 pp.)</p> <p>Nikola Halfarov&aacute;, Terezie Gajdo&scaron;ov&aacute;, Lenka M&aacute;lkov&aacute;, Nikol Taufrov&aacute;, Nela Koci&aacute;nov&aacute; (4th year, FPF SLU) in the subject of prof. Du&scaron;an Katu&scaron;č&aacute;k Digitization II. they prepared ca. 80 s. of GT transcriptions from various historical prints from the 18th and 19th centuries written in Czech (Svabian).</p> <p>The model has ID78289 SLOVAK Supermodel print&amp;typewriter (SSPT1) we used 542 pages in Ground Truth quality (GT 37897 lines and 200697 words). 59 pages for validation of the new model (Validation set). repositories of the State Scientific Library in Ostrava, the Slovak National Library in Martin, from the Manuscriptorium repository, from the State Archive in Bansk&aacute; Bystrica and from the Library of Matej Bel University in Bansk&aacute; Bystrica.<br>We consider the samples themselves to be very important for further use and refinement of the model, other researchers will get an idea of the similar or different writing of their own documents that they want to transcribe.</p> <p>In model ID78289 SLOVAK Supermodel print&amp;typewriter (SSPT1) the values Train set: 1.00% and Validation set: 1.00% were achieved. So it means the "accuracy" of the transcription is 99%.</p> <p>The creation of the SSM1 model on the Transkribus server took 21 hours and 52 minutes. The creation process was set to 100 cycles and ended after 100 cycles (epochs).<br>The SSPT1 model is the first attempt in Slovakia and the Czech Republic to create an aggregated tool through which it would be possible to automatically make available certain types of printed and typewritten documents that are similar to the fonts used in its creation.<br>In the event of an impossibility, we consider the ID78289 SLOVAK Supermodel print&amp;typewriter (SSPT1) to be the definitive universal model for the transcription of historical prints and typewriters of Slovak provenance of all types and periods. The variety of fonts and styles is diverse, and the creation of an optimal aggregate model is a task for other researchers and enthusiasts in the years to come.<br>However, we believe that our first aggregated SSPT1 model may be necessary for the automatic transcription of other analog documents.<br>The research team plans to make available datasets within the sustainability of the project in 2024-2028 prioritized for research and educational purposes for institutions and researchers who will want to contribute to the model of historical and new documents in West Slavic languages, respectively. languages of Slovak and Bohemian origin. Copyright: CC BY-NC-SA.<br>Of course, such automatic transcription will not immediately bring satisfactory results. However, it can be "rough" to gradually automatically transcribe additional pages, manually correct them to GT status, and then use larger GT datasets to refine a new model based on our SSPT1. After the creation of hundreds and thousands of pages of GT, it will be possible to proceed with the creation of further generations of new models based on SSPT1. Development could continue for printing and typewriting with models of new generations SSPT2, SSPT3 etc.</p> <p>The challenge for researchers is also the development and creation of a new aged supermodel, which would include both manuscripts and prints and typescripts. This Slovak supermodel could be shared within the Transkribus community of experts and included in some big Transkribus Community supermodel, etc.</p>

opencc-by-4.0May 2024View details →
zenodo48/100

PIBE project- Experimental characterization of stall noise in static and dynamic regimes using a NACA 63(3)418 airfoil

<p>Dynamic stall noise is one of the potential sources of amplitude modulations associated with wind turbine noise. This phenomenon is related to the periodic separation and reattachment of the boundary layer on the wind turbine blade suction side during its rotation. Within the framework of the PIBE project (Predicting the Impact of Wind Turbine Noise - <a href="https://www.anr-pibe.com/en">https://www.anr-pibe.com/en</a>), experiments were conducted in the anechoic wind tunnel of the &Eacute;cole Centrale de Lyon in order to characterize stall noise on a pitching airfoil in both static and dynamic conditions.</p> <p>In version 1.0.0 of the database, <span>data from the second campaign using an instrumented NACA63(3)418 airfoil in static and dynamic conditions are provided. The static data can be found in the file static_data_NACA63418.h5 that contains:</span></p> <ol> <li>static wall pressure data : lift and pressure coefficients;</li> <li>dynamic wall pressure data : Power Spectral Density (PSD) of fluctuating wall pressure;</li> <li>far-field acoustic data : Power Spectral Density (PSD) of acoustic pressure.</li> </ol> <p><span>The structure of the file is described in Tree_structure_static_data.pdf. To read the HDF5 file, the Matlab scripts given in read_HDF5_NACA63418_static_Matlab.zip can be used.</span></p> <p><span>The dynamic data can be found in the file dynamic_data_NACA63418.h5 that contains:</span></p> <ol> <li><span>static wall pressure data : phase-averaged lift coefficients;</span></li> <li><span>dynamic wall pressure data : phase-averaged spectrograms of fluctuating wall pressure;</span></li> <li><span>far-field acoustic data : phase-averaged spectrograms of acoustic pressure.</span></li> </ol> <p><span>The structure of the file is described in Tree_structure_dynamic_data.pdf. To read the HDF5 file, the Matlab scripts given in read_HDF5_NACA63418_dynamic_Matlab.zip can be used. Only the results for a mean angle of attack of 15&deg; and an amplitude of 15&deg; are provided in this file.</span></p>

opencc-by-4.0Jun 2024View details →
zenodo48/100

OECD Digitalization Dataset ODDEA Project

<p>Dataset collected and processed as part of the ODDEA (Overcoming Digital Divide Between Europe and Southeast Asia) EU research project (<em>Project ID: HORIZON MSCA-SE 101086381)</em>.The dataset consists of four OECD databases: Broadband and Telecommunication Database (23 indicators), The ICT Access and Usage by Households Database, The ICT Access and Usage by Individuals Database (106 indicators for households and individuals) , The ICT Usage by Business Database (59 indicators). The data are collected in Excel files (3) and csv file (1). They cover a period of 2012 to 2023 (if available) for OECD countries (40).</p>

opencc-by-4.0Jun 2024View details →
zenodo48/100

Project Tycho Level 2 data: Counts of multiple diseases reported in UNITED STATES OF AMERICA, 1888-2014

Project Tycho data include counts of infectious disease cases or deaths per time interval. A count is equivalent to a data point.<p></p><p>Project Tycho level 2 version 1.1.0 data include data counts that have been filtered from the raw data to render standardized data that can be used immediately for analysis. All level 2 data were originally reported in a consistent format and have not been transformed into a standard format by Project Tycho staff, except for smallpox records that included repeated counts for the same location and week, but sometimes with different numbers. These duplicate smallpox records have been averaged into one count for each location and week. Level 2 data include counts for a wide variety of diseases and locations for varying time periods. Because we removed data in an inconsistent format from level 2 data, counts may be missing for certain diseases, locations, or years. For the most complete collection of standardized data, we encourage users to use Project Tycho version 2.0 datasets.</p><p>More detailed methods and additional information about the origin of Projec Tycho level 2 version 1.1.0 data can be found in our original publication in the New England Journal of Medicine: <a href="http://www.nejm.org/doi/full/10.1056/NEJMms1215400">http://www.nejm.org/doi/full/10.1056/NEJMms1215400</a></p><p>Level 2 version 1.1.0 data is represented in a CSV file with 11 columns:</p><ul><li>epi_week: a six digit number that represents the year and epidemiological week for which disease cases or deaths were reported (yyyyww)</li><li>country: a two digit country abbreviation, only including "US" in version 1.1.0</li><li>state: the two digit postal code state abbreviation that represents the state for which a count has been reported</li><li>loc: the name of a state or city for which a count has been reported, capitalized</li><li>loc_type: the type of location (STATE or CITY) for which a count has been reported</li><li>disease: the disease for which a count has been reported, in all capitals</li><li>event: an indicator representing the disease outcome reported, including "CASES" or "DEATHS"</li><li>number: the reported number of cases or deaths</li><li>from_date: the start date of the time interval for which a count was reported, as yyyy-mm-dd</li><li>to_date: the end date of the time interval for which a count was reported, as yyyy-mm-dd</li><li>url: the URL of the source document from which the count was obtained</li></ul><p></p>

opencc-by-4.0Mar 2018View details →
zenodo48/100

Project Tycho Level 1 data: Counts of multiple diseases reported in UNITED STATES OF AMERICA, 1916-2011

<p>Project Tycho data include counts of infectious disease cases or deaths per time interval. A count is equivalent to a data point. Project Tycho level 1 data include data counts that have been standardized for a specific, published, analysis. Standardization of level 1 data included representing various types of data counts into a common format and excluding data counts that are not required for the intended analysis. In addition, external data such as population data may have been integrated with disease data to derive rates or for other applications.</p><p>Version 1.0.0 of level 1 data includes counts at the state level for smallpox, polio, measles, mumps, rubella, hepatitis A, and whooping cough and at the city level for diphtheria. The time period of data varies per disease somewhere between 1916 and 2011. This version includes cases as well as incidence rates per 100,000 population based on historical population estimates. These data have been used by investigators at the University of Pittsburgh to estimate the impact of vaccination programs in the United States, published in the New England Journal of Medicine: <a href="http://www.nejm.org/doi/full/10.1056/NEJMms1215400">http://www.nejm.org/doi/full/10.1056/NEJMms1215400</a>. See this paper for additional methods and detail about the origin of level 1 version 1.0.0 data.</p><p>Level 1 version 1.0.0 data is represented in a CSV file with 7 columns:</p><ul><li>epi_week: a six digit number that represents the year and epidemiological week for which disease cases or deaths were reported (yyyyww)</li><li>state: the two digit postal code state abbreviation that represents the state for which a count has been reported</li><li>loc: the name of a state or city for which a count has been reported, capitalized</li><li>loc_type: the type of location (STATE or CITY) for which a count has been reported</li><li>disease: the disease for which a count has been reported: HEPATITIS A, MEASLES, MUMPS, PERTUSSIS, POLIO, RUBELLA, SMALLPOX, or DIPHTHERIA</li><li>cases: the number of cases reported for the specified disease, epidemiological week, and location</li><li>incidence_per_100000: the number of cases per 100,000 people, computed using historical population counts for cities and states as reported by the US Census Bureau</li></ul><p></p>

opencc-by-4.0Mar 2018View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record