Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
6,025
datasets available to search
ShareScore release 0.9.0
Dataset results
6,025 results for “Science of science”
Peer Production in Citizen Science: A Community-Centered Approach on the Example of Personal Science - Supplementary Materials
<p>This folder contains supplementary materials for the PhD thesis “Peer Production in Citizen Science: A Community-Centered Approach on the Example of Personal Science” by Katharina Kloppenborg. This project has been led at Université Paris Cité / Inserm U1284 from 2020 to 2023 and was supervised by Bastian Greshake Tzovaras and Ariel Lindner. </p>
Co-design of a citizen science study: unlocking the potential of eDNA for volunteer freshwater monitoring
<ol> <li>Citizen science is increasingly being promoted as a means of gathering more data to help inform the management of ecosystems. Involving the participants in the design of data collection activities is a form of co-design often proposed by those calling for a translational ecology.</li> <li>In addition, novel monitoring approaches have the potential to improve the quality of data collected by citizen scientists. We explored the potential of environmental DNA (eDNA) for vertebrate (mainly fish) species monitoring through a co-designed catchment monitoring strategy. </li> <li>Having been introduced to the potential of eDNA, citizen scientists designed and executed an eDNA-based survey of a small chalk stream catchment to explore questions of concern.</li> <li>The eDNA monitoring approach provided data about fish and other vertebrate diversity in the catchment which would have otherwise required sampling approaches difficult for citizen scientists. These data give a preliminary answer to some of the citizen scientists' priority questions and are comparable to fish data collected through traditional electrofishing surveys.</li> <li>Recommendations are offered for co-design and the use of novel research techniques by citizen scientists.</li> </ol>
Le Journal of Open Humanities Data (JOHD) : enjeux et défis dans la publication de data papers pour les sciences humaines et sociales (SHS)
<p>This repository hosts the pre-print version of the paper and the data used for the following publication:</p> <p>Marongiu, Paola; Pedrazzini, Nilo; Ribary, Marton; McGillivray, Barbara (forthcoming). "<strong>Le <em>Journal of Open Humanities Data</em> (<em>JOHD</em>) : enjeux et défis dans la publication de <em>data papers</em> pour les sciences humaines et sociales (SHS)</strong>". In Schopfel, Joachim and Kosmopoulos, Christine (eds.) Publier, partager, réutiliser les données de la recherche: les data papers et leurs enjeux. Lille: Presses Universitaires du Septentrion.</p> <p>The file "johd_publications.xlsx" contains the data used for the section "Publier des<em> data papers</em> pour les SHS : l’expérience du<em> JOHD</em>". This includes all relevant information about the papers published by the Journal of Open Humanities data. The title of the published paper; the DOI; the date of publication; the special collection in which the paper was published (if any); the keywords associated with the paper; the names of the authors.</p> <p>The zip folder "datapapers_catalyseurs.zip" contains the data used for the section "Les <em>data papers</em> : des catalyseurs de la réutilisation des données ?". This includes information about the papers published by the Journal of Open Humanities data and their corresponding datasets (downloads, views, tweets, citations and Altmetric).</p>
Net-zero 1.5 °C sectorial pathways for G20 countries: energy and emissions data to inform science-based decarbonization targets
<p><span>This data for global, regional (EU-27), and country-specific (G20 member countries) energy and emission pathways required to achieve a defined carbon budget of under 450 Gt/CO2, developed to limit the mean global temperature rise to 1.5°C, over 50% likelihood. The data were calculated with the 1.5°C sectorial pathways of the One Earth Climate Model—an integrated energy assessment model devised at the University of Technology Sydney (UTS). </span></p> <p><span>The data consist of the following six zip-folder datasets (refer to Section 2 for an explanation of the data):</span></p> <p><span>1. </span><span>Appendix folder: Each file contains one worksheet, which summarizes the overall 1.5°C scenario.</span></p> <p><span>2. </span><span>Sector folder (XLSX): Each file contains one worksheet, which summarizes the industry sectors analysed.</span></p> <p><span>3. </span><span>Sector folder (CSV): The data contained are the same as those described in point 2.</span></p> <p><span>4. </span><span>Sector emissions folder: Each file contains one worksheet, which summarizes the total annual emissions for each industry sector.</span></p> <p><span>5. </span><span>Scope emissions folder (XLSX): Each file contains one worksheet, which summarizes the total annual emissions for each industry sector—with the additional specificity of emission scope. </span></p> <p><span>6. </span><span>Scope emissions folder (CSV): The data contained are the same as those described in point 5.</span></p>
The data of "A comparison of citation-based clustering and topic modeling for science mapping"
<p>These files consist of the data used in "A comparison of citation-based clustering and topic modeling for science mapping". </p> <p> </p>
Education in the Anthropocene: Assessing planetary health science standards in the US
The environmental crises defining the Anthropocene demand ubiquitous mitigation efforts, met with collective support. Yet, disengagement and disbelief surrounding planetary health threats are pervasive, especially in the United States (US). This skepticism may be influenced by inadequate education addressing the scope and urgency of the planetary health crisis. We analyzed current K-12 science standards related to planetary health throughout the US, assessing their quality and potential predictors of variation. While planetary health education varies widely across the US with respect to the presence and depth of terms, most science standards neglected to convey these concepts with a sense of urgency. Furthermore, state/territory political affiliation and primary GDP contributor were each predictive of the quality of planetary health education. We propose that a nation-wide science standard could fully address the urgency of the planetary health crisis and prevent political bias from influencing the breadth and depth of concepts covered.
Protéger les enquêtés, mais à quelles conditions ? Anonymiser des données d'enquêtes en sociologie et en science politique
<p>This dataset represents the analysis of a sample of 681 journal articles, published from 2012 to 2023. The dataset is used in the related by paper by the same authors.</p>
Supplementary information to "What does ChatGPT know about natural science and engineering?"
<p>This Excel workbook contains the survey data and data analysis from the manuscript "What does ChatGPT know about natural science and engineering?" by Schulze Balhorn et al.</p>
Datasets used in "Evidential Deep Learning: Enhancing Predictive Uncertainty Estimation for Earth System Science Applications"
<p>The precipitation type (p-type) dataset (ptype.parquet) comprises observational weather reports sourced from the Meteorological Phenomena Identification Near the Ground (mPING) project, combined with corresponding numerical weather prediction data from the NOAA Rapid Refresh (RAP) model. These crowd-sourced mPING reports offer precipitation type labels (rain, snow, sleet, and freezing rain) across North America, while the RAP model provides atmospheric data, including temperature, humidity, and wind profiles, on pressure levels.</p> <p> </p> <p>The RAP data covers the contiguous United States (CONUS) from 2015 to 2022 on an hourly 13km grid. The mPING observations are matched to the nearest RAP grid cell and hour, allowing the two data sources to be merged into a labeled dataset suitable for classification tasks. </p> <p> </p> <p>The surface layer flux dataset (surface_layer.csv) contains high-frequency meteorological observations spanning from 2013 to 2015, collected at the Cabauw Experimental Site in the Netherlands. It includes measurements of various variables such as temperature, humidity, wind, radiation, and soil moisture, recorded every 10 minutes. The target output encompasses friction velocity, sensible heat, and latent heat.</p> <p><br> The code used for processing the datasets and training neural network models is available in the Miles-Guess repository (<a href="https://github.com/ai2es/miles-guess">https://github.com/ai2es/miles-guess</a>).</p>
Dataset for: Fifty years of research on questionable research practices in science: Quantitative analysis of co-citation patterns
<p>Questionable research practices (QRPs) have been the focus of the scientific community amid greater scrutiny and evidence highlighting issues with replicability across many fields of science. To capture the most impactful publications and the main thematic domains in the literature on QRPs, this study uses a document co-citation analysis. The analysis was conducted on a sample of 341 documents that covered the past 50 years of research in QRPs. Nine major thematic clusters emerged. Statistical reporting and statistical power emerged as key areas of research, where systemic-level factors in how research is conducted are consistently raised as the precipitating factors for QRPs. There is also an encouraging shift in the focus of research into open science practices designed to address engagement in QRPs. Such a shift is indicative of the growing momentum of the open science movement, and more research can be conducted on how these practices are employed on the ground and how their uptake by researchers can be further promoted. However, the results suggest that, while pre-registration and registered reports receive the most research interest, less attention has been paid to other open science practices (e.g., data and methods sharing).</p>
Imptox Public Workshop Micro- Nanoplastics & Human Health: Insights from Science and Research (pt2)
<p>Experts gather at the Imptox workshop to discuss the emerging understanding of micro- and nanoplastics' impact on human health. On Friday, March 24th, 2023, the successful Imptox public workshop attracted a large audience discussing the potential health impacts of micro- and nanoplastics with experts from all over Europe. The Imptox project has received funding from the EU’s H2020 framework program for research and innovation under grant agreement n. 965173</p>
Imptox Public Workshop Micro- Nanoplastics & Human Health: Insights from Science and Research (pt3)
<p>Experts gather at the Imptox workshop to discuss the emerging understanding of micro- and nanoplastics' impact on human health. On Friday, March 24th, 2023, the successful Imptox public workshop attracted a large audience discussing the potential health impacts of micro- and nanoplastics with experts from all over Europe. The Imptox project has received funding from the EU's H2020 framework program for research and innovation under grant agreement n. 965173. </p>
Plant science corpus
<p>The plant science corpus consists of the titles and abstracts of plant science articles in PubMed published prior to 2021 with a small number of 2021 records due to modification of records. The columns are:</p><ul><li>Index: integer index serving as identifier</li><li>PMID: PubMed identifier</li><li>Date: Publication date</li><li>Journal: journal where the article was published</li><li>Title: Title of the article</li><li>Abstract: Abstract of the article</li><li>Corpus: Title and abstract combined</li><li>Text classification score: plant science record prediction model score</li><li>Preprocessed corpus: Corpus after lower-casing, stop word removal, removal of non-alphanumeric and non-white space characters, lemmitisation</li><li>Topic: index of topics after topic modeling</li></ul>
European Organization for Nuclear Research and "Marek-Lars Kruusen's technology and science"
<p>"Marek-Lars Kruusen's technology and science" is a company primarily engaged in scientific research on wormholes and technology development. Marek-Lars Kruusen collaborates with the European Organization for Nuclear Research (CERN). This collaboration involves the publication of scientific papers and preprints on a platform managed by CERN. For example, preprints can be found on Zenodo, which is hosted at CERN.</p> <p>The European Organization for Nuclear Research, known as CERN, is an intergovernmental organization that operates the largest particle physics laboratory in the world. Established in 1954, it is based in a suburb of Geneva, on the France–Switzerland border. It comprises 23 member states. Israel, admitted in 2013, is the only non-European full member. CERN is an official United Nations General Assembly observer.</p> <p>The acronym CERN is also used to refer to the laboratory; in 2019, it had 2,660 scientific, technical, and administrative staff members, and hosted about 12,400 users from institutions in more than 70 countries. In 2016, CERN generated 49 petabytes of data.</p> <p>CERN's main function is to provide the particle accelerators and other infrastructure needed for high-energy physics research – consequently, numerous experiments have been constructed at CERN through international collaborations. CERN is the site of the Large Hadron Collider (LHC), the world's largest and highest-energy particle collider. The main site at Meyrin hosts a large computing facility, which is primarily used to store and analyze data from experiments, as well as simulate events. As researchers require remote access to these facilities, the lab has historically been a major wide area network hub. CERN is also the birthplace of the World Wide Web.</p> <p>Since its foundation by 12 members in 1954, CERN regularly accepted new members. All new members have remained in the organization continuously since their accession, except Spain and Yugoslavia. Spain first joined CERN in 1961, withdrew in 1969, and rejoined in 1983. Yugoslavia was a founding member of CERN but quit in 1961. Of the 23 members, Israel joined CERN as a full member on 6 January 2014, becoming the first (and currently only) non-European full member.</p> <p>The Open Science movement focuses on making scientific research openly accessible and on creating knowledge through open tools and processes. Open access, open data, open source software and hardware, open licenses, digital preservation and reproducible research are primary components of open science and areas in which CERN has been working towards since its formation.</p>
Multiplier Event 7 (E7) : Citizen Science for Librarians: open access resources from a teaching/learning perspective
<p>Event: Piliečių mokslas bibliotekininkams: atvirosios prieigos ištekliai mokymo(si) perspektyvoje (Citizen Science for Librarians: open access resources from a<br>teaching/learning perspective)<br>Date: 24.05.2024<br>Language: Lithuanian<br>Target group: Lithuanian academic and public library staff</p> <p>LIBOCS Multiplier Event number 7 </p> <p>Objectives of the event:<br>1. present the open access collection of resources on citizen science (PR6A1)<br>2. to present Citizen Science toolkit for Librarians (PR6A2) and its potential or use in<br>teaching and learning<br>3. to demonstrate all the project results on Zenodo platform</p> <p><strong>Project acronym</strong>: LibOCS<br><strong>Full project title</strong>: University libraries strengthening the academia-society connection through citizen science in the Baltics</p> <p>This project is funded under the Erasmus+ KA2 Strategic Partnerships program. Project Number: 2021-1-EE01-KA220-HED-000031125</p>
Austrian Science Fund (FWF) Publication Cost Data 2013 - 2018
<p>The data set consists of the FWF publication cost data sets of the years 2013 to 2018 which are publicly available on Zenodo: <a href="https://zenodo.org/communities/fwf/?page=1&size=20">https://zenodo.org/communities/fwf/?page=1&size=20</a></p> <p>The data set includes payments for publications funded by the FWF programmes:</p> <p>Peer-reviewed Publications: <a href="https://www.fwf.ac.at/en/research-funding/fwf-programmes/peer-reviewed-publications/">https://www.fwf.ac.at/en/research-funding/fwf-programmes/peer-reviewed-publications/</a></p> <p>and Stand-Alone Publications: <a href="https://www.fwf.ac.at/en/research-funding/fwf-programmes/stand-alone-publications/">https://www.fwf.ac.at/en/research-funding/fwf-programmes/stand-alone-publications/</a></p> <p> </p>
Leveraging Implementation Science to Increase Access to Trauma Treatment for Incarcerated Drug Users
ClinicalTrials.gov study NCT04007666. IPD Sharing: YES. Countries: 1. Publications: 1.
MAPS Trial: Matrix And Platinum Science
ClinicalTrials.gov study NCT00396981. IPD Sharing: Not stated. Countries: 11. Publications: 2.
Behavioural Science Messages in Breast Cancer Screening
ClinicalTrials.gov study NCT05395871. IPD Sharing: NO. Countries: 1. Publications: 6.
Citizen Science to Promote Sustained Physical Activity in Low-Income Communities
ClinicalTrials.gov study NCT03041415. IPD Sharing: YES. Countries: 1. Publications: 2.
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.