Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
87
datasets available to search
ShareScore release 0.9.0
Dataset results
87 results for “data documentation”
Data set from the 6G-SANDBOX platforms benchmarking and assessment for different documented trial networks - Oulu facility
<p>This data set covers core network and end-to-end measurements at the Oulu facility, part of the 6G-SANDBOX infrastructure.</p>
Data set from the 6G-SANDBOX platforms benchmarking and assessment for different documented trial networks - Malaga facility
<p>This data set covers core network and end-to-end measurements at the Malaga facility, part of the 6G-SANDBOX infrastructure.</p>
Data set from the 6G-SANDBOX platforms benchmarking and assessment for different documented trial networks - Berlin facility
<p>This data set covers core network and end-to-end measurements at the Berlin facility, part of the 6G-SANDBOX infrastructure.</p>
Data set from the 6G-SANDBOX platforms benchmarking and assessment for different documented trial networks - Athens facility
<p>This data set covers core network and end-to-end measurements at the Athens facility, part of the 6G-SANDBOX infrastructure.</p>
Data from: Open notes sounds great, but will a provider's documentation change?
<p><strong>Background</strong>: The effects of shared clinical notes on patients, care partners, and clinicians ("open notes") were first studied as a demonstration project in 2010. Since then, multiple studies have shown clinicians agree shared progress notes are beneficial to patients, and patients and care partners report benefits from reading notes. To determine if implementing open notes at a hematology/oncology practice changed providers' documentation style, we assessed the length and readability of clinicians' notes before and after open notes implementation at an academic medical center in Boston, MA.</p> <p><strong>Methods</strong>: We analyzed 143,888 notes from 60 hematology/oncology clinicians before and after the open notes debut at Beth Israel Deaconess Medical Center, from January 1, 2012, to September 1, 2016. We measured the providers' (medical doctor/nurse practitioner) documentation styles by analyzing character length, the number of addenda, note entry mode (dictated vs. typed) and note readability. Measurements used five different readability formulas and were assessed on notes written before and after the introduction of open notes on November 25, 2013.</p> <p><strong>Results</strong>: After the introduction of open notes, the mean length of progress notes increased from 6,174 characters to 6,648 characters (P<0.001), and the mean character length of the "assessment and plan" (A&P) increased from 1,435 characters to 1,597 characters (P<0.001). The Average Grade Level Readability of progress notes decreased from 11.50 to 11.33, and overall readability improved by 0.17 (P=0.01). There were no statistically significant changes in the length or readability of "Initial Notes" or Letters, inter-doctor communication, nor in the modality of the recording of any kind of note.</p> <p><strong>Conclusions</strong>: After the implementation of open notes, progress notes and A&P sections became both longer and easier to read. This suggests clinician documenters may be responding to the perceived pressures of a transparent medical records environment.</p>
Data for Seed-driven Document Ranking for Systematic Reviews: A Reproducibility Study
<p>Data for Seed-driven Document Ranking for Systematic Reviews: A Reproducibility Study</p>
Documentation data for the Python Apecosm package
<p>This dataset contains NEMO/Pisces and Apecosm output files that are used in the documentation of the Apecosm python package.</p>
Data from: Documenting the progressions of secondary eyewall formations
Open the record for dataset details and reuse information.
Data from: Characterizing population structure and documenting rapid loss of genetic diversity in Chiricahua Leopard Frogs (Lithobates chiricahuensis) with high throughput microsatellite genotyping
Open the record for dataset details and reuse information.
Data from: Open notes sounds great, but will a provider’s documentation change?
Open the record for dataset details and reuse information.
Dataset for "Understanding Performance Concerns in the API Documentation of Data Science Libraries"
<p>Dataset for the manuscript "Understanding Performance Concerns in the API Documentation of Data Science Libraries", including the results of knowledge classification, consistency analysis, and evolution analysis on the API documentation data.</p>
Survey Data Set Part 2 - Attitudes Towards Videos as a Documentation Option for Communication in Requirements Engineering
<p>In 2017, I conducted an online survey to explore software professionals' attitudes towards videos as a documentation option for communication in requirements engineering. The survey covered the following topics:</p> <ul> <li>Demographics</li> <li>Attitude towards videos as a medium in RE including its strengths, weaknesses, opportunities, and threats</li> <li>Current production and use of videos in RE, respectively the obstacles that prevent the production and use of videos</li> </ul> <p>64 out of 106 software professionals from industry and academia completed the survey. The survey was implemented in LimeSurvey and distributed across several communication channels such as LinkedIn, ResearchGate, Twitter, and a mailing list of a German RE professionals group.</p> <p>This dataset includes the following files:</p> <ul> <li>"Raw and analyzed data.xlsx" contains the raw and analyzed survey responses which are anonymized <ul> <li>This data includes <em>demographics </em>and <em>video production and use</em>.</li> <li>The data on <em>attitudes</em> are included in: <a href="https://zenodo.org/record/3245770">Survey Data Set Part 1 - Attitudes Towards Videos as a Documentation Option for Communication in Requirements Engineering</a>.</li> </ul> </li> <li>"Survey - Offline version.docx" contains the questions and possible answers of the survey</li> <li>"Survey - Offline version.pdf" contains the questions and possible answers of the survey</li> </ul> <p>This survey was designed, conducted, and analyzed by Oliver Karras (<a href="https://twitter.com/KarrasOliver">@KarrasOliver</a>)</p>
Blue-Action D5.13 Supplementary data, documentation and figures
<p>This archives supplementary information for D5.13 Evaluation of the polar lows forecast system (related to Case Study 3: <a href="https://doi.org/10.5281/zenodo.4294460">https://doi.org/10.5281/zenodo.4294460</a>). The data produced are uploaded. Descriptions of methods and scripts are given in:</p> <p>To_create_Marine_Cold_Air_Outbreaks_MCAO_data_from_ERA_Interim.pdf</p> <p>Analysing_anomalies_of_a_few_weather_variables_associated_with_MCAO_event_durations.pdf</p> <p>Further analyses that are not presented in D5.13 are also provided:</p> <p>kingetal_2020_weatherpatternsassociatedwithmcaos.pdf</p> <p>king_2019jul_slidesforwp1group.pdf</p>
Dispatches from the neighborhood watch: using citizen science and field survey data to document color morph frequency in space and time
<p>Heritable color polymorphisms have a long history of study in evolutionary biology, though they are less frequently examined today than in the past. These systems, where multiple discrete, visually identifiable color phenotypes co-occur in the same population, are valuable for tracking evolutionary change and ascertaining the relative importance of different evolutionary mechanisms. Here, we use a combination of citizen science data and field surveys in the Great Lakes region of North America to identify patterns of color morph frequencies in the eastern gray squirrel (<i>Sciurus carolinensis</i>). Using over 68,000 individual squirrel records from both large and small spatial scales, we identify the following patterns: (1) the melanistic (black) phenotype is often localized but nonetheless widespread throughout the Great Lakes region, occurring in all states and provinces sampled. (2) In Ohio, where intensive surveys were performed, there is a weak but significantly positive association between color morph frequency and geographic proximity of populations. Nonetheless, even nearby populations often had radically different frequencies of the melanistic morph, which ranged from 0 to 96%. These patterns were mosaic rather than clinal. (3) In the Wooster, Ohio population, which had over eight years of continuous data on color morph frequency representing nearly 40,000 records, we found that the frequency of the melanistic morph increased gradually over time on some survey routes but decreased or did not change over time on others. These differences were statistically significant and occurred at very small spatial scales (on the order of hundreds of meters). Together, these patterns are suggestive of genetic drift as an important mechanism of evolutionary change in this system. We argue that studies of color polymorphism are still quite valuable in advancing our understanding of fundamental evolutionary processes, especially when coupled with the growing availability of data from citizen science efforts.</p>
Data from: The challenge of accurately documenting bee species richness in agroecosystems: bee diversity in eastern apple orchards
Bees are important pollinators of agricultural crops, and bee diversity has been shown to be closely associated with pollination, a valuable ecosystem service. Higher functional diversity and species richness of bees have been shown to lead to higher crop yield. Bees simultaneously represent a mega-diverse taxon that is extremely challenging to sample thoroughly and an important group to understand because of pollination services. We sampled bees visiting apple blossoms in 28 orchards over 6 years. We used species rarefaction analyses to test for the completeness of sampling and the relationship between species richness and sampling effort, orchard size, and percent agriculture in the surrounding landscape. We performed more than 190 h of sampling, collecting 11,219 specimens representing 104 species. Despite the sampling intensity, we captured <75% of expected species richness at more than half of the sites. For most of these, the variation in bee community composition between years was greater than among sites. Species richness was influenced by percent agriculture, orchard size, and sampling effort, but we found no factors explaining the difference between observed and expected species richness. Competition between honeybees and wild bees did not appear to be a factor, as we found no correlation between honeybee and wild bee abundance. Our study shows that the pollinator fauna of agroecosystems can be diverse and challenging to thoroughly sample. We demonstrate that there is high temporal variation in community composition and that sites vary widely in the sampling effort required to fully describe their diversity. In order to maximize pollination services provided by wild bee species, we must first accurately estimate species richness. For researchers interested in providing this estimate, we recommend multiyear studies and rarefaction analyses to quantify the gap between observed and expected species richness.
Data from: A new dimension in documenting new species: high-detail imaging for myriapod taxonomy and first 3D cybertype of a new millipede species (Diplopoda, Julida, Julidae)
We review the state-of-the-art approaches currently applied in myriapod taxonomy, and we describe, for the first time, a new species of millipede (Ommatoiulus avatar n. sp., family Julidae) using high-resolution X-ray microtomography (microCT) as a substantive adjunct to traditional morphological examination. We present 3D models of the holotype and paratype specimens and discuss the potential of this non-destructive technique in documenting new species of millipedes and other organisms. The microCT data have been uploaded to an open repository (Dryad) to serve as the first actual millipede cybertypes to be published.
Data from: A low-cost solution for documenting distribution and abundance of endangered marine fauna and impacts from fisheries
Fisheries bycatch is a widespread and serious issue that leads to declines of many important and threatened marine species. However, documenting the distribution, abundance, population trends and threats to sparse populations of marine species is often beyond the capacity of developing countries because such work is complex, time consuming and often extremely expensive. We have developed a flexible tool to document spatial distribution and population trends for dugongs and other marine species in the form of an interview questionnaire supported by a structured data upload sheet and a comprehensive project manual. Recognising the effort invested in getting interviewers to remote locations, the questionnaire is comprehensive, but low cost. The questionnaire has already been deployed in 18 countries across the Indo-Pacific region. Project teams spent an average of USD 5,000 per country and obtained large data sets on dugong distribution, trends, catch and bycatch, and threat overlaps. Findings indicated that >50% of respondents had never seen dugongs and that 20% had seen a single dugong in their lifetimes despite living and fishing in areas of known or suspected dugong habitat, suggesting that dugongs occured in low numbers. Only 3% of respondents had seen mother and calf pairs, indicative of low reproductive output. Dugong hunting was still common in several countries. Gillnets and hook and line were the most common fishing gears, with the greatest mortality caused by gillnets. The questionnaire has also been used to study manatees in the Caribbean, coastal cetaceans along the eastern Gulf of Thailand and western Peninsular Malaysia, and river dolphins in Peru. This questionnaire is a powerful tool for studying distribution and relative abundance for marine species and fishery pressures, and determining potential conservation hotspot areas. We provide the questionnaire and supporting documents for open-access use by the scientific and conservation communities.
Data from: A metacalibrated time-tree documents the early rise of flowering plant phylogenetic diversity
The establishment of modern terrestrial life is indissociable from angiosperm evolution. While available molecular clock estimates of angiosperm age range from the Paleozoic to the Late Cretaceous, the fossil record is consistent with angiosperm diversification in the Early Cretaceous. The time-frame of angiosperm evolution is here estimated using a sample representing 87% of families and sequences of five plastid and nuclear markers, implementing penalized likelihood and Bayesian relaxed clocks. A literature-based review of the palaeontological record yielded calibrations for 137 phylogenetic nodes. The angiosperm crown age was bound within a confidence interval calculated with a method that considers the fossil record of the group. An Early Cretaceous crown angiosperm age was estimated with high confidence. Magnoliidae, Monocotyledoneae and Eudicotyledoneae diversified synchronously 135–130 million yr ago (Ma); Pentapetalae is 126–121 Ma; and Rosidae (123–115 Ma) preceded Asteridae (119–110 Ma). Family stem ages are continuously distributed between c. 140 and 20 Ma. This time-frame documents an early phylogenetic proliferation that led to the establishment of major angiosperm lineages, and the origin of over half of extant families, in the Cretaceous. While substantial amounts of angiosperm morphological and functional diversity have deep evolutionary roots, extant species richness was probably acquired later.
Raw documents of data (eggshell project)
<p>Raw documents of data for eggshell project</p>
Deep sequencing data for document titled: Rolling circle RNA synthesis catalysed by RNA
<p>RNA-catalysed RNA replication is widely considered a key step in the emergence of life's first genetic system. However, RNA replication can be impeded by the extraordinary stability of duplex RNA products, which must be dissociated for re-initiation of the next replication cycle. Here we have explored rolling circle synthesis (RCS) as a potential solution to this strand separation problem. RCS on small circular RNAs - as indicated by molecular dynamics simulations - induces a progressive build-up of conformational strain with destabilisation of nascent strand 5' and 3' ends. At the same time, we observe sustained RCS by a triplet polymerase ribozyme on small circular RNAs over multiple orbits with strand displacement yielding concatemeric RNA products. Furthermore, we show RCS of a circular Hammerhead ribozyme capable of self-cleavage and re-circularisation. Thus, all steps of a viroid-like RNA replication pathway can be catalysed by RNA alone. Our results have implications for the emergence of RNA replication and for understanding the potential of RNA to support complex genetic processes.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.