Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
209
datasets available to search
ShareScore release 0.9.0
Dataset results
209 results for “anonymization”
Anonymization of Clinical Data From Pseudonymized Databases Collected as Part of Previus Clinical Trials on Multiple Myeloma
ClinicalTrials.gov study NCT07283224. IPD Sharing: Not stated. Countries: 1. Publications: 0.
Anonymous Testing of Pathology Specimens for BRCA Mutations in Ashkenazi Jewish Individuals Who Have Cancer
ClinicalTrials.gov study NCT00588263. IPD Sharing: Not stated. Countries: 1. Publications: 0.
Assessment of Preoperative Anxiety and Evaluation of Calming Interventions: Anonymous Patient Survey
ClinicalTrials.gov study NCT07291466. IPD Sharing: Not stated. Countries: 1. Publications: 0.
Using Anonymous Data from a Digital Tool for Medical History-Taking to Improve Healthcare Services
ClinicalTrials.gov study NCT06761105. IPD Sharing: UNDECIDED. Countries: 1. Publications: 0.
Anonymous
<p>Anonymous</p>
Anonymous Submission
<p>Trained models and data sets for anonymous submission</p>
Anonymized_transcriptions_of_focus_groups
<p><a href="https://zenodo.org/api/files/365e2d0c-a161-47ba-b5cf-d351f309dffb/Anonymized_transcriptions_of_focus_groups.docx">Anonymized_transcriptions_of_focus_groups.docx</a></p>
Collection of Biological and Environmental Samples and Clinical Data From Anonymous Adult Men and Women for Quality Control and Methods Development and Evaluation (ASCA)
ClinicalTrials.gov study NCT02940015. IPD Sharing: Not stated. Countries: 0. Publications: 0.
Anonymized version of the dataset repository
<div> </div> <div>This is the replication data for the scientific paper titled "Processed food intake assortativity in the personal networks of East European older adults." For details on how to use the data files, please consider the "Supplementary Material" file and the manuscript where the context of the study and the variables of interest are presented. This is an anonymized version of the dataset repository. We prepared this anonymized version only for the double-blind peer review process conducted by the Health & Place journal. The reviewers can access the data set and the corresponding code for evaluation and replication purposes. The un-anonymized version of the dataset is already freely available and will be disclosed upon completion of the paper's review. </div>
10M anonymized Twitter human and bot profiles correlated with 2022 Russo-Ukrainian War discussion.
Open the record for dataset details and reuse information.
Program and Operational Characteristics of Syringe Services Programs—Data from the National Survey of Syringe Services Programs (Anonymized), 2022
<p><span>The National Survey of Syringe Services Programs (NSSSP) is an activity under the </span><span>Centers for Disease Control and Prevention’s</span><span> (CDC’s) PS22-2208 “</span><span>Support a national network of Syringe Services Programs (SSPs) and oversee implementation and use of an annual survey of SSPs”,</span><span> Component 1. The goal of the NSSSP is to collect information on SSP characteristics, services and referrals offered, people served, funding, community relationships, and operational successes and challenges annually to ensure comprehensive services are being provided to meet the needs of people who use drugs (PWUD). These data will be used to monitor national SSP coverage, document the capacity and needs of SSPs, quantify the public health impact of SSPs, and identify disparities in areas with low coverage relative to a high burden of health consequences from the syndemic of drug use, HIV, and hepatitis C virus (HCV). These data are not intended to be used to determine the effectiveness of SSPs. The NSSSP was reviewed by the RTI Institutional Review Board (IRB) and while it does not collect any individual-level data or personally identifying information, the NSSSP does contain sensitive organization level information which requires protection.</span></p>
Anonymized interview transcripts from the study "Information Scientists' Motivations for Research Data Sharing and Reuse"
<p>These are eleven transcripts of the interviews conducted in the context of the study “Information Scientists’ Motivations for Research Data Sharing and Reuse”.</p> <p>Interview partners are information science researchers who work with data and shared (or not) their research data and/or reused (or not) the data from third parties. Study sample included six women and five men working either at a German university and/or a research performing organization with working experience as researchers from six to over 40 years. Nine persons have a PhD degree, two persons are the PhD candidates. One researcher belongs to the Net Generation (i.e. born approx. between 1946 and 1960), three researchers – to Generation X (i.e. born approx. 1960 and 1980), seven researchers – to Generation Y (i.e. born approx. between 1980 and 1996).</p> <p>Data collection took place between October, 13 and November, 24 2022. The interviews were conducted in German, online and video-recorded. The shortest conversation took 43 minutes and the longest – 63 minutes. With participants’ consent, all interviews were transcribed using f4x Automatic Speech Recognition. The automatically created transcripts were then manually checked and corrected if needed.</p> <p>The interview transcripts are anonymized by replacing original words or phrases with more generic words or phrases. The replaced texts are marked in blue and taken in brackets. The access to the anonymized transcripts will be granted on request with concrete agreement on use.</p> <p>The <a href="https://doi.org/10.5281/zenodo.8230979">interview guide template</a> used for conducting the interviews is available. For more information about study design and results, please visit the <a href="https://doi.org/10.5281/zenodo.8186518">project page</a> and/or read the article "Information Scientists' Motivations for Research Data Sharing and Reuse" by Shutsko and Stock (2023).</p>
Anonymized
<p>V2</p>
AI in Higher Education: Does Not Help, Might Hurt (Anonymized Dataset)
Open the record for dataset details and reuse information.
Program and Operational Characteristics of Syringe Services Programs—Data from the National Survey of Syringe Services Programs (Anonymized), 2023
<p>The National Survey of Syringe Services Programs (NSSSP) is an activity under the Centers for Disease Control and Prevention’s (CDC’s) PS22-2208 “Support a national network of Syringe Services Programs (SSPs) and oversee implementation and use of an annual survey of SSPs”, Component 1. The goal of the NSSSP is to collect information on SSP characteristics, services and referrals offered, people served, funding, community relationships, and operational successes and challenges annually to ensure comprehensive services are being provided to meet the needs of people who use drugs (PWUD). These data will be used to monitor national SSP coverage, document the capacity and needs of SSPs, quantify the public health impact of SSPs, and identify disparities in areas with low coverage relative to a high burden of health consequences from the syndemic of drug use, HIV, and hepatitis C virus (HCV). These data are not intended to be used to determine the effectiveness of SSPs. The NSSSP was reviewed by the RTI Institutional Review Board (IRB) and while it does not collect any individual-level data or personally identifying information, the NSSSP does contain sensitive organization level information which requires protection.</p>
Posts from a brazilian anonymous imageboard
<p>This dataset contains text with toxic content and hate speech.</p> <p>A set of discussion threads published in a brazilian anonymous imageboard, a 4chan-style discussion forum. This dataset includes 158,280 user posts in 4,539 threads published between 18 dec. 2016 and 19 jan. 2017. The data was collected through a web scraper developed for this purpose, which gattered textual content and published date from the posts. Images where not collected due to possible ilegal content.</p> <p>The data was used in the paper "A manifestação da masculinidade tóxica em um fórum de internet anônimo brasileiro".</p>
Anonymized Dataset for "How Difficult Can It Be? Assessing Workload Components and Predictors in Introductory CS Projects"
<p>Paper title: How Difficult Can It Be? Assessing Workload Components and Predictors in Introductory CS Projects</p> <p>This is a replication package using the anonymized dataset that we have collected including 161 entries of NASA-TLX surveys and the associated student source code metrics. Please cite the paper properly if you plan on using any parts of the dataset or the replication package.</p> <p> <br> <br> The following preprocessing steps proceed the work in this file: <br> 1- Collecting survey data and filtering incomplete entries, anonymizing data by replacing names with numbers. <br> 2- Calculating TLX and adding it as a feature to the data. <br> 3- Processing the code associated with each entry to calculate: loc, functions/methods, classes, style errors. <br> 4- Processing submission system logs to calculate number of commits/submissions. <br> <br> </p> <p><br> Dataset: anonymous-complete-preprocessed-TLX.csv <br> key: anonymous-dataset-key.txt</p>
Anonymous
<p>dd</p>
Anonymous record hosting dataset for our submission
Open the record for dataset details and reuse information.
Anonymous
<p>Anonymous</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.