Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
1,026
datasets available to search
ShareScore release 0.9.0
Dataset results
1,026 results for “Linked data”
Dataset linking to the publication "An assessment of data sources, data quality and changes in national forest monitoring capacities in the Global Forest Resources Assessment 2005–2020"
<p>This dataset links to the study “An assessment of data sources, data quality and changes in national forest monitoring capacities in the Global Forest Resources Assessment 2005–2020”. This study is published in the journal “Environmental Research Letters” which can be found at <a href="https://iopscience.iop.org/article/10.1088/1748-9326/abd81b">https://iopscience.iop.org/article/10.1088/1748-9326/abd81b</a>. The dataset contains two files, one csv file, and one shape file. The two files contain the same data to meet the different users' needs. The dataset contains variables for assessing national forest monitoring data sources i.e., RS and/or NFI. Separate indicators namely 'Use of RS', and 'Use of NFI' were used to analyze the two data sources (RS and NFI). The description of each variable for these two indicators contained in the dataset is given in the Table below.</p> <table> <caption><strong>The description of the variables in the datase</strong>t <strong>for country capacity assessment</strong></caption> <tbody> <tr> <td><strong>Variables Name</strong></td> <td><strong>Description of the variables</strong></td> </tr> <tr> <td>Country</td> <td>Country</td> </tr> <tr> <td>ISO_A3_CODE</td> <td>ISO A3 Code for country</td> </tr> <tr> <td>ADM0_CODE</td> <td>ADMO Code for country</td> </tr> <tr> <td>CONTINENT</td> <td>Continent</td> </tr> <tr> <td>Region</td> <td>Region</td> </tr> <tr> <td>RSInd_05</td> <td>Use of remote sensing (RS) for forest area (change) monitoring 2005 Indicator</td> </tr> <tr> <td>RSSc_05</td> <td>Use of RS for forest area (change) monitoring 2005 Score</td> </tr> <tr> <td>RSInd _10</td> <td>Use of RS for forest area (change) monitoring 2010 Indicator</td> </tr> <tr> <td>RSSc _10</td> <td>Use of RS for forest area (change) monitoring 2010 Score</td> </tr> <tr> <td>RSInd_15</td> <td>Use of RS for forest area (change) monitoring 2015 Indicator</td> </tr> <tr> <td>RSSc _15</td> <td>Use of RS for forest area (change) monitoring 2015 Score</td> </tr> <tr> <td>RSInd_20</td> <td>Use of RS for forest area (change) monitoring 2020 Indicator</td> </tr> <tr> <td>RSSc _20</td> <td>Use of RS for forest area (change) monitoring 2020 Score</td> </tr> <tr> <td>DRS05_20</td> <td>Difference ‘use of RS’ 2005-2020</td> </tr> <tr> <td>NFIInd_05</td> <td>Use of national forest inventories (NFI) for forest monitoring 2005 Indicator</td> </tr> <tr> <td>NFISc_05</td> <td>Use of NFI for forest monitoring 2005 Score</td> </tr> <tr> <td>NFIInd _10</td> <td>Use of NFI for forest monitoring 2010 Indicator</td> </tr> <tr> <td>NFISc _10</td> <td>Use of NFI for forest monitoring 2010 Score</td> </tr> <tr> <td>NFIInd_15</td> <td>Use of NFI for forest monitoring 2015 Indicator</td> </tr> <tr> <td>NFISc _15</td> <td>Use of NFI for forest monitoring 2015 Score</td> </tr> <tr> <td>NFIInd_20</td> <td>Use of NFI for forest monitoring 2020 Indicator</td> </tr> <tr> <td>NFISc _20</td> <td>Use of NFI for forest monitoring 2020 Score</td> </tr> <tr> <td>DNFI05_20</td> <td>Difference ‘Use of NFI’ 2005-2020</td> </tr> </tbody> </table> <p>Indicators and Scores in the above Table for showing the use of RS and NFI data for forest monitoring in Figure 1 (1a, 1b, and 2a, 2b) are related in the following way.</p> <table> <caption><strong>The indicator values and scores of the country capacity assessment</strong></caption> <tbody> <tr> <td><strong>Indicator</strong></td> <td><strong>Score</strong></td> </tr> <tr> <td>Low</td> <td>0</td> </tr> <tr> <td>Limited</td> <td>1</td> </tr> <tr> <td>Intermediate</td> <td>2</td> </tr> <tr> <td>Good</td> <td>3</td> </tr> <tr> <td>Very Good</td> <td>4</td> </tr> </tbody> </table> <p>The capacity changes from 2005 to 2020 in Figure 1 (1c & 2c) are related in the following way.</p> <table> <caption><strong>The indicator values and levels for country capacity changes</strong></caption> <tbody> <tr> <td><strong>Capacity change values</strong></td> <td><strong>Capacity change levels</strong></td> </tr> <tr> <td>1,2,3,4</td> <td>Increase</td> </tr> <tr> <td>0</td> <td>No change</td> </tr> <tr> <td>-1,-2,-3,-4</td> <td>Decrease</td> </tr> </tbody> </table> <p> </p>
Linked Open Data Management Services: A Comparison
<p>Thanks to a variety of software services, it has never been easier to produce, manage and publish Linked Open Data. But until now, there has been a lack of an accessible overview to help researchers make the right choice for their use case. This dataset release will be regularly updated to reflect the latest data published in a comparison table developed in Google Sheets [1]. The comparison table includes the most commonly used LOD management software tools from NFDI4Culture to illustrate what functionalities and features a service should offer for the long-term management of FAIR research data, including:</p> <ul> <li>ConedaKOR</li> <li>LinkedDataHub</li> <li>Metaphacts</li> <li>Omeka S</li> <li>ResearchSpace</li> <li>Vitro</li> <li>Wikibase</li> <li>WissKI</li> </ul> <p>The table presents two views based on a comparison system of categories developed iteratively during workshops with expert users and developers from the respective tool communities. First, a short overview with field values coming from controlled vocabularies and multiple-choice options; and a second sheet allowing for more descriptive free text additions. The table and corresponding dataset releases for each view mode are designed to provide a well-founded basis for evaluation when deciding on a LOD management service. The Google Sheet table will remain open to collaboration and community contribution, as well as updates with new data and potentially new tools, whereas the datasets released here are meant to provide stable reference points with version control.</p> <p>The research for the comparison table was first presented as a paper at DHd2023, Open Humanities – Open Culture,<strong> </strong>13-17.03.2023, Trier and Luxembourg [2].</p> <p>[1] Non-editing access is available here: <a href="http://docs.google.com/spreadsheets/d/1FNU8857JwUNFXmXAW16lgpjLq5TkgBUuafqZF-yo8_I/edit?usp=share_link">docs.google.com/spreadsheets/d/1FNU8857JwUNFXmXAW16lgpjLq5TkgBUuafqZF-yo8_I/edit?usp=share_link</a> To get editing access contact the authors.</p> <p>[2] Full paper will be made available open access in the conference proceedings.</p>
Ports, Past and Present Blog (perma.cc link tabular data)
<p>A list of blog and calendar posts from the Ports, Past and Present project written on Wordpress. This tabular data is a modified output from the <a href="https://perma.cc/">perma.cc</a> folder containing a series of WARC records captured on the 26th of June, 2023.</p>
Ha-SingleMoleculeLab's data for publication: Linking folding dynamics and function of SAM/SAH riboswitches at the single molecule level
<p>This upload is the raw data that support our findings sent to review on Nucleic Acids Research, corresponding to each individual figure. The paper title is " Linking folding dynamics and function of SAM/SAH riboswitches at the single molecule level".</p>
Online supplementary data linked to the publication "Aubenas-les-Alpes (S-E France). Part III – Last and final part of the mammalian assemblage with some comments on the palaeoenvironment and palaeobiogeography" doi:10.1016/j.annpal.2019.03.001
<p>Online supplementary appendix including the list of Oligocene localities and associated faunal lists compared to Aubenas-les-Alpes, and the size estimation of the non-predatory species for the construction of Fig.10.</p>
10 Women of Digital Humanities: An Analysis of Linked Open Data
<p>This is a spreadsheet of linked open data, including tweets of the 10 female scholars randomly chosen for this project. This is an experimental study, and was prepared for my own learning. Presentation prepared for a Masters seminar in Information Science at uOttawa in Winter 2020. </p> <p>Linked Open Data in the Humanities was taught by Prof. Constance Crompton.</p> <p>10 Women of Digital Humanities: An Analysis of Linked Open Data</p> <p>tags: dh, digital humanities, feminist dh, computational analysis, Voyant, word clouds, vizualization, digital identifiers, open scholarship</p>
Data from: Linking pollen foraging of megachilid bees to their nest bacterial microbiota
<p>Solitary bees build their nests by modifying the interior of natural cavities and they provision them with food by importing collected pollen. As a result, the microbiota of the solitary bee nests may be highly dependent on introduced materials. In order to investigate how the collected pollen is associated with the nest microbiota, we used metabarcoding of the ITS2 rDNA and the 16S rDNA to simultaneously characterize the pollen composition and the bacterial communities of 100 solitary bee nest chambers belonging to seven megachilid species. We found a weak correlation between bacterial and pollen alpha-diversity and significant associations between the composition of pollen and that of the nest microbiota, contributing to the understanding of the link between foraging and bacteria acquisition for solitary bees. Since solitary bees cannot establish bacterial transmission routes through eusociality, this link could be essential for obtaining bacterial symbionts for this group of valuable pollinators.</p>
Data from: Linking land use and the nutritional ecology of herbivores: a case study with the Senegalese locust
1) Access to high-quality food is a main driver of population dynamics. For herbivores protein and carbohydrates are key nutrients that are notoriously variable in plants and are affected by land use. However, few studies have linked foraging decisions and performance in the laboratory to the nutritional landscape available in the field. 2) <i>Oedaleus senegalensis</i> is a nonmodel locust, a grass-feeder, and the main pest of millet, a subsistence crop in the Sahel. In this study, we examined dietary preference and locust performance across a range of protein:carbohydrate ratios using the Geometric Framework methodology. We then applied a fitness landscape approach to visualize these results with the plant nutrient contents available across four land-use types: millet, groundnut, fallow, and grazed fields. Finally, we contrasted our results with locust distribution in the field. Several locust species (<i>O. senegalensis</i> included) exhibit density dependent color polymorphism thus we also reported individual coloration (brown or green). 3) We found that <i>O. senegalensis</i> preferred moderately carbohydrate-biased food 1:1.6 protein: carbohydrate ratio. All traits recorded (mass gain, development time, growth rate, molt success, and performance index) were best near that ratio and declined on either side presenting a "hump-shape". Fallow fields contained more plants, particularly grasses, that were both abundant and closer to the optimal protein:carbohydrate ratio recorded from the lab experiments. 4) When we surveyed <i>O. senegalensis</i> abundance and proportion, we found that they were more numerous in the fallow fields. Brown morph individuals, the ones associated with high density, were proportionally more abundant in fallow fields than green individuals. 5) Our study provides evidence that variation in nutritional landscapes—relative to an herbivore's optimal nutrient balance—is a key driver of herbivore population distribution and abundance, and can be used to predict bottom-up effects on herbivore species. protein:carbohydrate ratio recorded from the lab experiments.
Data and code for the paper Atmospheric Observations with E-band Microwave Links – Challenges and Opportunities
<p>Raw and preprocessed data for the paper Atmospheric Observations with E-band Microwave Links – Challenges and Opportunities accepted for publication in the journal Atmospheric Measurement Techniques.</p> <p>The dataset includes total losses (transmitted - received power levels) of commercial microwave links and observations of rainfall, air temperature, and air relative humidity. In addition, theoretical gaseous attenuation calculated from air temperature and relative humidity is provided.</p> <p>Data are stored in semicolon-delimited csv files. Time stamps are in UTC time in the format yyyy-mm-dd HH:MM:SS. Raw data contain not regular time series, preprocessed data contain regular time series at 1-min and 5-min temporal resolution. Metadata are stored in textfiles.</p> <p>Dataset contains also R code for processing and analyzing the data as presented in the paper Atmospheric Observations with E-band Microwave Links – Challenges and Opportunities. The code is in the form of R Markdown files and interactive html notebooks.</p>
Data for: Improving Kieker's Scalability by Employing Linked Read-Optimized and Write-Optimized NoSQL Storage
<p>We show, how polyglot persistence increase Kieker's scalability by employing separate read-optimized and write-optimized noSQL storage. For this purpose we extended Kieker to store its monitoring output in Apache Cassandra , which is a write-optimized wide-column noSQL database. For the analysis of the generated monitoring output we are using ElasticSearch, a read-optimized document store noSQL storage. We are interlinking read-optimized and write-optimized noSQL storage within our Regression Benchmarking Execution Environment (RBEE). To ensure scalability we are employing a container infrastructure. The mentioned noSQL storages and the linker are operating within one single Docker container which scales horizontally.</p> <p>For generating reference values, we instrumented a Java SE application with Kieker's file system writer and measured throughput and method's execution times. Consecutively, we instrumented the same Java SE application with our Apache Cassandra writer and measured throughput and method's execution time. Finally, we compared measurement results of Kieker's file system writer with the measurement results of our Apache Cassandra writer.</p>
The MIDI Linked Data Cloud dataset
<p>The study of music is highly interdisciplinary, and thus requires the combination of datasets from multiple musical domains, such as catalog metadata (authors, song titles, dates), industrial records (labels, producers, sales), and music notation (scores). While today an abundance of music metadata exists on the Linked Open Data cloud, datasets containing interoperable symbolic descriptions of music itself, i.e. music notation with note and instrument level information, are scarce. This is the MIDI Linked Data Cloud, a dataset that represents multiple collections of digital music in the MIDI standard format as Linked Data. At the time of writing, the dataset comprises 10,215,557,355 triples of 308,443 interconnected MIDI scores, and provides Web-compatible descriptions of their MIDI events.</p>
Data from: warming and top-down control of stage-structured prey: linking theory to patterns in natural systems
<p>Warming has broad and often nonlinear impacts on organismal physiology and traits, allowing it to impact species interactions like predation through a variety of pathways that may be difficult to predict. Predictions are commonly based on short-term experiments and models, and these studies often yield conflicting results depending on the environmental context, spatiotemporal scale, and the predator and prey species considered. Thus, the accuracy of predicted changes in interaction strength, and their importance to the broader ecosystems they take place in, remain unclear. Here, we attempted to link one such set of predictions generated using theory, modeling, and controlled experiments to patterns in the natural abundance of prey across a broad thermal gradient. To do so, we first predicted how warming will impact a stage-structured predator-prey interaction in riverine rock pools between Pantala spp. dragonfly nymph predators and Aedes atropalpus mosquito larval prey. We then described temperature variation across a set of hundreds of riverine rock pools (n = 775) and leveraged this natural gradient to look for evidence for or against our model's predictions. Our model's predictions suggested that warming should weaken predator control of mosquito larval prey by accelerating their development and shrinking the window of time that aquatic dragonfly nymphs could consume them in. This was consistent with data collected in rock pool ecosystems, where the negative effects of dragonfly nymph predators on mosquito larval abundance were weaker in warmer pools. Our findings provide additional evidence to substantiate our model-derived predictions, while emphasizing the importance of assessing similar predictions using natural gradients of temperature whenever possible.</p>
Data from: Seed preference is only weakly linked to seed-type-specific feeding performance in a songbird
<p>The dehusking of seeds by granivorous songbirds is a complex process that requires fast, coordinated and sensory-feedback-controlled movements of beak and tongue. Hence, efficient seed handling requires a high degree of sensorimotoric skill and behavioural flexibility, since seeds vary considerably in size, shape and husk structure. To deal with this variability, individuals might specialise on specific seed types, which could result in greater seed handling efficiency of the preferred seed type, but lower efficiency for other seed types. To test this, we assessed seed preferences of canaries (Serinus canaria) through food choice experiments and related these to data of feeding performance, seed handling skills and beak kinematics during feeding on small, spindle-shaped canary seeds and larger, spheroid-shaped hemp seeds. We found great variety in seed preferences among individuals: some had no clear preference, while others almost exclusively fed on hemp seeds, or even prioritized novel seed types (millet seed). Surprisingly, we only observed few and weak effects of seed preference on feeding efficiency. This suggests that either the ability to handle seeds efficiently can be readily applied across various seed types, or alternatively, it may indicate that achieving high levels of seed-specific handling skills does not require extensive practice.</p>
Data from: Metabarcoding analysis provides insight into the link between prey and plant intake in a large alpine cat carnivore, the snow leopard
<p>Species of the family Felidae (a group represented by cats) are thought to be obligate carnivores, specialized for hunting and consuming other animals. However, the detection of plants in the feces of felids raises questions about the role of plants in their diet. This is particularly true for the snow leopard (Panthera uncia), a big cat native to central and South Asia's high mountains. Our study aimed to comprehensively identify the prey and plants consumed by snow leopards as well as six other sympatric mammals. We applied DNA metabarcoding methods on 126 fecal samples collected from the Sarychat-Ertash Nature Reserve in Kyrgyzstan. We found that among the three most common plant families in snow leopard feces, Tamaricaceae (genus Myricaraia) was consumed often by snow leopards. The genus Myricaria frequently appeared in samples lacking any animal prey DNA, indicating that snow leopards might have consumed this plant especially when their digestive tracts were empty. We also observed a significant difference in plant composition between male and female snow leopards, and potentially between sampling seasons. We provide a comprehensive overview of the prey and plants detected in the feces of snow leopards and sympatric mammals. We believe our findings will help in formulating hypotheses and guiding future research to understand the adaptive significance of plant-eating behavior in felids and animal-plant relationships in the ecosystem.</p>
Data associated with the publication "Multi-decadal increase of forest burned area in Australia is linked to climate change"
<p>Data from various sources related to fires in Australian forests, associated with the publication "Multi-decadal increase of forest burned area in Australia is linked to climate change"</p>
Datasets for Linked Open Data Instance Level Analysis for Cultural Heritage
<p>This is the datasets used for Linked Open Data instant level quality analysis for cultural heritage (2020). 7Z and ZIP versions are available for both Excel 2006 and R 4.0.3. The compressed files include, Excel spreadsheets (.xlsx, .csv), VBA scripts (.bas), and R scripts (.r).</p> <p>Please read the full documentation in Linked_Open_Data_Instance_Level_Analysis_Procedure.pdf.</p>
Process Controls on Flood Seasonality in Brazil - link to data and plots
<p>This data set and plots of circular histograms accompany the paper "Process Controls on Flood Seasonality in Brazil" at Geophysical Research Letters (<a href="https://doi.org/10.1029/2021GL096754">https://doi.org/10.1029/2021GL096754</a>). In this paper, we investigate the relationship between the seasonality of floods, maximum annual rainfall, and maximum annual soil moisture data of 886 basins in Brazil for 1980-2015 to shed light on process controls of flood generation.</p>
Data from: Strong links between plant traits and microbial activities but different abiotic drivers in mountain grasslands
<p>This dataset contains data and code that support the results in Weil, S.-S., Martinez-Almoyna, C., Piton, G., Renaud, J., Boulangeat, L., Foulquier, A., ... & Thuiller, W. (2021) Strong links between plant traits and microbial activities but different abiotic drivers in mountain grasslands (accepted in Journal of Biogeography).</p> <p>We used an extensive plant-soil dataset that covers 14 elevational gradients (between 1500 and 2800 m of elevation) distributed over the whole French Alps to analyse the spatial co-dependencies between the plant and soil compartments. We ran a Graphical Lasso that extracts the direct and indirect linkages between plant functional composition, soil microbial activities, and environmental conditions (local climate and soil properties).</p> <p>Our main results are 1) that plant traits are tightly associated with microbial activities, the former being driven by climate and the latter by soil properties; 2) that the dominance of specific plant traits was more important than their diversity to determine plant-soil linkages; and 3) that soil microbes invested strongly in nutrient acquisition in sites with conservative plant traits and reduced organic matter quality.</p>
Integrated Statistical Indicators from Scottish Linked Open Government Data
<p>Integrated statistical indicators that were retrieved from the official Scottish data portal in order to facilitate the exploitation of Machine Learning methods in Open Government Data. Data include 60 statistical indicators from seven categories such as health and social care, housing, and crime and justice. The indicators refer to the 6,976 “2011 data zones” of Scotland, while the year of reference is 2015. Data are ready to be used by the research community, students, policy makers, and journalists and give rise to plenty of social, business, and research scenarios that can be solved using Machine Learning technologies and methods.</p>
Data published in manuscript "Highest methane concentrations in an Arctic River linked to local terrestrial inputs"
<p>This data is published in the manuscript:</p> <p>Castro-Morales, K., Canning, A., Arzberger, S., Overholt, W.A., Küsel, K., Kolle, O., Göckede, M., Zimov, N. and Körtzinger, A. (2022). Highest methane concentrations in an Arctic River linked to local terrestrial inputs. <em>Biogeosciences.</em> XX, XXX-XXX. https://doi.org/10.5194/bg-XX-XXX-2022.</p> <p>The data contains the water properties, the dissolved gas concentrations and flux densities at 1-min resolution corresponding to two transects in the Kolyma River main channel and two tributaries (Ambolikha and Leonid). The data was collected between 15 and 17 June, 2019.<strong> </strong></p> <p>This folder contains four data files and the file "README_Data_access_Castro-Morales_etal_CH4_Kolyma_River.txt" provides more details on the data.</p> <p> </p> <p> </p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.