Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
753
datasets available to search
ShareScore release 0.9.0
Dataset results
753 results for “metrics”
Group and individual social network metrics are robust to changes in resource distribution in experimental populations of forked fungus beetles
<p>Social interactions drive many important ecological and evolutionary processes. It is therefore essential to understand the intrinsic and extrinsic factors that underlie social patterns. A central tenet of the field of behavioral ecology is the expectation that the distribution of resources shapes patterns of social interactions.</p> <p>We combined experimental manipulations with social network analyses to ask how patterns of resource distribution influence complex social interactions.</p> <p>We experimentally manipulated the distribution of an essential food and reproductive resource in semi-natural populations of forked fungus beetles (Bolitotherus cornutus). We aggregated resources into discrete clumps in half of the populations and evenly dispersed resources in the other half. We then observed social interactions between individually marked beetles. Half-way through the experiment, we reversed the resource distribution in each population, allowing us to control any demographic or behavioral differences between our experimental populations. At the end of the experiment, we compared individual and group social network characteristics between the two resource distribution treatments.</p> <p>We found a statistically significant but quantitatively small effect of resource distribution on individual social network position and detected no effect on group social network structure. Individual connectivity (individual strength) and individual cliquishness (local clustering coefficient) increased in environments with clumped resources, but this difference explained very little of the variance in individual social network position. Individual centrality (individual betweenness) and measures of overall social structure (network density, average shortest path length, and global clustering coefficient) did not differ between environments with dramatically different distributions of resources.</p> <p>Our results illustrate that the resource environment, despite being fundamental to our understanding of social systems, does not always play a central role in shaping social interactions. Instead, our results suggests that sex differences and temporally fluctuating environmental conditions may be more important in determining patterns of social interactions.</p>
Datasample - Social Media Analytics and Metrics of Facebook Performance of Libraries, Archives and Museums
<p>The current dataset describes Facebook pages performance for 220 Libraries, Archives and Museums from all over the world. The performance is measured through 9 different social media metrics. That is, number of posts, link-posts, picture-posts, video-posts, total reactions, comments and shares, number of reactions, comments per post and reactions per post. The data harvesting process has been conducted through the use of FanPageKarma API. The gathered metrics and their values depict the performance for each Facebook page in a time-period of 30 days.</p>
LiDAR metrics generated from Airborne Laser Scanning (ALS) data across the Netherlands
<p>This data repository contains the LiDAR metrics generated from country-wide Airborne Laser Scanning (ALS) data from the Netherlands. The LiDAR metrics (10-meter resolution) are derived from AHN3 using <a href="https://laserfarm.readthedocs.io/en/latest/">Laserfarm</a> workflow. Raw point cloud data can be downloaded <a href="https://app.pdok.nl/ahn3-downloadpage/">here</a>. </p>
Microservice Security Metrics: Dataset
<p>This is the dataset provided for replicability for the article "Microservice Security Metrics for Secure Communication, Identity Management, and Observability". It provides the code needed to replicate the study in the article, as well as the model data set of 10 system models and 20 variants of those models.</p> <p> </p> <p>The abstract of the article is:</p> <p> </p> <p>Microservice architectures are increasingly being used to develop application systems. Despite many guidelines and best practices being published, architecting microservice systems for security is challenging. Reasons are the size and complexity of microservice systems, their polyglot nature, and the demand for the continuous evolution of these systems. In this context, to manually validate that security architecture tactics are employed as intended throughout the system is a time-consuming and error-prone task. In this article, we present an approach to avoid such manual validation before each continuous evolution step in a microservice system, which we demonstrate using three widely used categories of security tactics: secure communication, identity management, and observability. Our approach is based on a review of existing security guidelines, the gray literature, and the scientific literature, from which we derived Architectural Design Decisions (ADDs) with the found security tactics as decision options. In our approach, we propose novel detectors to detect these decision options automatically and formally defined metrics to measure the conformance of a system to the different options of the ADDs. We apply the approach on a case study data set of 10 open source microservice systems, plus another 20 variants of these systems, for which we manually inspected the source code for security tactics. We demonstrate and assess the validity and appropriateness of our metrics by performing an assessment of their conformance to the ADDs in our systems' dataset through statistical methods.</p>
Metrics for continuous delivery: a rapid review
<p>This is a replication package as part of the paper submitted for ESEM'22 (ACM / IEEE International<br> Symposium on Empirical Software Engineering and Measurement (ESEM)</p>
Galaxy Training Data for "Evaluating and ranking a set of pathways based on multiple metrics"
<pre>This dataset provides the inputs needed for the Galaxy Pathway Analysis workflow training tutorial (<a href="https://galaxy-synbiocad.org">https://galaxy-synbiocad.org</a>). This workflow asseses the performance of predicted pathways by computing 4 criteria (target product flux, thermodynamic feasibility, pathway length, and enzyme availability). A score inform the user about the best candidate pathways to produce a compound of interest. The generated output is a collection of scored and ranked heterologous pathways. The content of the dataset is as follows: - A set of pathways provided in the SBML format (Systems Biology Markup Language) to be ranked, modeling heterologous pathways such as those outputted by the RetroSynthesis workflow (<a href="https://galaxy-synbiocad.org">https://galaxy-synbiocad.org</a>). - The GEM (Genome-scale metabolic models) which is a formalized representation of the metabolism of the host organism (the model is E. coli iML1515), provided in the SBML format.</pre>
Opt2Q Calibrations -- Gelman Rubin Convergence Metrics
<p>Gelman Rubin Convergence metrics for calibrations detailed at</p> <p><a href="https://doi.org/10.5281/zenodo.4768814">https://doi.org/10.5281/zenodo.4768814</a><br> <a href="https://doi.org/10.5281/zenodo.4768842">https://doi.org/10.5281/zenodo.4768842</a><br> <a href="https://doi.org/10.5281/zenodo.4768806">https://doi.org/10.5281/zenodo.4768806</a><br> <a href="https://doi.org/10.5281/zenodo.4768370">https://doi.org/10.5281/zenodo.4768370</a></p>
Hypothetical landscapes to evaluate connectivity metrics of protected area networks.
<p>This repo contains the raw datasets (as GIS shapefiles) useful to evaluate connectivity metrics of protected area networks. Please suggest if additional landscapes could be added that would be useful to evaluate an additional class or characteristic of protected area networks. They were created using Google Earth Engine script: <strong><a href="https://code.earthengine.google.com/d1a8dfa3202ac8b4e55657bd3b5a1160">https://code.earthengine.google.com/d1a8dfa3202ac8b4e55657bd3b5a1160</a>.</strong></p> <p>Two shapefiles are provided: (1) ProNet_connectivity_library_L1_26pa -- this contains polygons that represent the size and shape of protected areas (PAs); (2) ProNet_connectivity_library_L1_26pae -- this contains polylines that represent "edges" that do not represent any protected area but denotes that two PAs are connected. Note that these landscapes are fictitious, and represented at the global origin (i.e. 0.0 degrees latitude and 0.0 degrees longitude) -- and are quite small so zooming in will be required to see them in GIS software.</p>
FAIR Metrics SKOS Mapping
<p>SKOS mapping between four FAIR metrics/indicators, based on "Best Effort".</p> <p>Created within the EOSC Synergy project task 3.3.</p> <ul> <li>RDA FAIR data maturity model indicators</li> <li>FAIRsFAIR Data Object Assessment Metrics</li> <li>FAIR Maturity Indicators from FAIRsharing</li> <li>FAIR Enough data maturity indicators</li> </ul>
A Curated Solidity Smart Contracts Repository of Metrics and Vulnerabilities
<p>SmarthER provides the dataset related to the full-paper accepted to PROMISE 2024 (<a href="https://promiseconf.github.io/2024/index.html" rel="nofollow">https://promiseconf.github.io/2024/index.html</a>) <strong>"A Curated Solidity Smart Contracts Repository of Metrics and Vulnerability"</strong>.</p> <p>Authored by: Giacomo Ibba, Sabrina Aufiero, Rumyana Neykova, Silvia Bartolucci, Roberto Tonelli, Marco Ortu, Giuseppe Destefanis</p> <p>This repository aims to collect a significant sample of smart contracts with associated vulnerability reports, and traditional software metrics extracted from each smart contract. The repository contains:</p> <ul> <li>Smart contracts source code.</li> <li>The vulnerability report was built with Slither for each contract.</li> <li>Traditional software metrics extracted from each contract.</li> </ul>
Survey Protocol - A Metrics suite for End-to-End Microservice Test Coverage
<p>These survey is performed and reported in a paper titled "A Metrics suite for End-to-End Microservice Test Coverage".</p>
Sharpness metric results and Instron method script for 'Raw Material Sharpness and Lithic Patterns: An Analysis of Holocene Susitna River Basin, Central Alaska' MPhil dissertation
<p>Data for appendix B in 'Raw Material Sharpness and Lithic Patterns: An Analysis of Holocene Susitna River Basin, Central Alaska' MPhil dissertation. This data set Includes initial and dulled sharpness metric data from the Instron® experiment for each flake. It also includes the Instron® bluehill method program script.</p>
Рис. 6. МоΔеΛирование экоΛогических ниш коΛораΔского жука ΔΛя ΔаΛьневосточного, европейского и североамериканского ареаΛов метоΔом метрического Δвухмерного шкаΛирования с применением коэффициента Жаккара Fig. 6. Models of ecological niches of the Colorado potato beetle for the Far Eastern, European, and North-American habitats (metric multidimensional scaling, Jaccard index) in Comparative characterization of the ecology of native (Henosepilachna vigintioctomaculata) and invasive (Leptinoatrsa decemlineata) species under the conditions of the monsoon climate in the southern part of the Russian Far East
Рис. 6. МоΔеΛирование экоΛогических ниш коΛораΔского жука ΔΛя ΔаΛьневосточного, европейского и североамериканского ареаΛов метоΔом метрического Δвухмерного шкаΛирования с применением коэффициента Жаккара Fig. 6. Models of ecological niches of the Colorado potato beetle for the Far Eastern, European, and North-American habitats (metric multidimensional scaling, Jaccard index)
Beyond Throughput: a 4G LTE Dataset with Channel and Context Metrics
<p>The following provides a 4G trace dataset composed of client-side cellular key performance indicators (KPIs) collected from two major Irish mobile operators, across different mobility patterns (static, pedestrian, car, tram and train). The 4G trace dataset contains 135 traces, with an average duration of fifteen minutes per trace, with viewable throughput ranging from 0 to 173 Mbit/s at a granularity of one sample per second. Our traces are generated from a well-known non-rooted Android network monitoring application, G-NetTrack Pro. This tool enables capturing various channel related KPIs, context-related metrics, downlink and uplink throughput, and also cell-related information.</p> <p>To supplement our real-time 4G production network dataset, we also provide a synthetic dataset generated from a large-scale 4G ns-3 simulation that includes one hundred users randomly scattered across a seven-cell cluster. The purpose of this dataset is to provide additional information (such as competing metrics for users connected to the same cell), thus providing otherwise unavailable information about the eNodeB environment and scheduling principle, to end user. In addition to this dataset, we also provide the code and context information to allow other researchers to generate their own synthetic datasets.</p>
Research data, sources and documents for thesis on Exploring Complexity Metrics for Artifact-Centric Business Process Models
<p>Research data, sources and documents for thesis on Exploring Complexity Metrics for Artifact-Centric Business Process Models This repository contains the supplemental material for the <a href="https://pqdtopen.proquest.com/pubnum/10759956.html">thesis "Exploring Complexity Metrics for Artifact-Centric Business Process Models" by Marin, Mike A., Ph.D., University of South Africa (South Africa), 2017.</a></p>
Stable Modeling on Resource Usage Parameters of MapReduce Application-Figure 4. Statistical Metrics distribution of model on RIO as response of Terasort application
<p>Figure 4 shows the statistical metrics distribution of regression model on read rate as the response of Terasort application. The filled triangle point-up indicates the minimum stable sampling time for statistical metrics. The top-half of figure 4 shows the residual standard error (RSE) distribution as training data size increase. The remaining half is for the distribution of R2.</p>
Stable Modeling on Resource Usage Parameters of MapReduce Application-Figure 8. Minimum sample time of statistical metrics of MapReduce applications
<p>Figure 8 presents the minimum sampling time distribution of statistic metrics which ensures the stable modeling. Overall, the minimum sampling time of statistic metrics is smaller than sampling time of estimated coefficients. For different applications, a time-consuming application like Terasort needs the largest sampling time to tend to be stable. The Pi application shows the smallest minimum sampling time to reach stability.</p>
Data for "Sounding out Ecoacoustic Metrics: Avian species richness is predicted by acoustic indices in temperate but not tropical habitats"
<p>This deposit contains the data for the paper <strong>A Multi-habitat, Comparative Evaluation of Ecoacoustic Indices for Biodiversity Monitoring: Acoustic Indices Predict Avian Species Richness in Temperate but not Tropical Habitats. (Ecological Indicators) </strong>The dataset contains a series of 1 min wav files recorded across UK and Ecuadorian habitats. Each one has 26 acoustic indices calculated on it, and a full list of avian species and abundances and GPS data for each sample site.</p> <p>Abstract</p> <p>Affordable, autonomous recording devices facilitate large scale acoustic monitoring and Rapid Acoustic Survey is emerging as a cost-effective approach to ecological monitoring; the success of the approach rests on the development of computational methods by which biodiversity metrics can be automatically derived from remotely collected audio data. Dozens of indices have been proposed to date, but systematic validation against classical, in situ diversity measures. This study conducted the most comprehensive comparative evaluation to date of the relationship between avian species diversity and a suite of acoustic indices across a wide range of ecological conditions. Acoustic surveys were carried out across habitat gradients in temperate and tropical biomes. Baseline avian species richness and subjective multi-taxa biophonic density estimates were established through aural counting by expert ornithologists. 26 acoustic indices were calculated and compared to observed variations in species diversity. Five acoustic diversity indices (Bioacoustic Index, Acoustic Diversity Index, Acoustic Evenness Index, Acoustic Entropy, and the Normalised Difference Sound Index) were assessed as well as three simple acoustic descriptors (root-mean-square, spectral centroid and zero-crossing rate). Highly significant correlations, of up to 65%, between acoustic indices and avian species richness were observed across temperate habitats, supporting the use of automated acoustic indices in biodiversity monitoring where a single vocal taxon dominates. Significant, weaker correlations were observed in neotropical habitats which host multiple non-avian vocalizing species. Multivariate classification analyses suggest that AIs also track observed differences in habitat-dependent community composition and that each habitat has a distinct soundscape. Multivariate analyses of the relative predictive power of AIs show that compound indices are more powerful predictors of avian species richness than any single index and simple descriptors contribute to predicting avian diversity in multi-taxa tropical environments. Our results support the use of community level acoustic indices as a proxy for species richness and point to the potential for tracking of habitat-dependent changes in community composition. Recommendations for the design of compound indices for multi-taxa community composition appraisal are put forward, with consideration for the requirements of next generation, low power remote monitoring networks.</p> <p> </p> <p><strong>Sampling Methods (extract from paper)</strong></p> <p>Acoustic surveys were carried out along a gradient of habitat degradation (1 forested, 2 regenerating forest and 3 agricultural land) in South East (SE) England and North Western (NW) Ecuador. The six sites (UK1, UK2, UK3, EC1, EC2, EC3) were sampled consecutively from May 6th - Aug 25th 2015.</p> <p>All UK sites were in the county of Sussex, in SE England, an area of weald clays (Fig. 2, left) and included ancient woodland (UK1), regenerating farmland with patches of woodland (UK2) and a downland barley farm (UK3).1 min mono audio recordings made every 15 minutes at three different habitats in the UK</p> <p>Ten day acoustic surveys were carried out consecutively at each study site using 15 Wildlife Acoustics Song Meter audio field recorders. Sampling points were arranged in a grid at a minimum distance of 200 m to minimise pseudo replication (the sound of most species being attenuated over this distance in all biomes). Altitudinal range of sample points across sites was minimised in order to prevent introduction of extraneous, confounding gradients (UK varied between 10 m – 50 m and Ecuador 130 m – 390 m). Recording schedules captured 1 min every 15 min around the clock for 10 days at each site, resulting in 960 recordings at each of 15 sample points for 3 habitat types in 2 different climates (86,400 1 minute recordings in total). Data across the 15 sample points was pooled; inter-site variation was not explored in the current analyses. In the UK 3½ hours of each dawn chorus was sampled starting at 1 hour before sunrise. This range was determined to capture the onset, progression and peak of the dawn chorus, creating a temporal gradient. The equatorial dawn chorus is more compact and was sampled for 2¼ hours starting 15 mins before sunrise, capturing a comparable chorus onset and peak.</p> <p> </p> <p> </p>
Spatial patterns of uncertainty in climate exposure metrics for North America at 1km resolution
<p>The data provided below represents the degree of uncertainty or variation between 8 individual general circulation models (GCM) for three metrics commonly used to assess the intensity of exposure to climate change. The three exposure metrics (forward and backward <a href="https://adaptwest.databasin.org/pages/adaptwest-velocitywna">climatic velocity</a> and <a href="https://adaptwest.databasin.org/pages/climatic-dissimilarity">local climatic dissimilarity</a>) were calculated based on the first two principal components (PC) scores derived from <a href="https://adaptwest.databasin.org/pages/climatic-dissimilarity">11 different climate variables</a>. Frameworks and heuristics supporting climate adaptation for conservation often rely on projections of climate change or climate exposure. However, projections of climate change vary among alternative GCM outputs, different emissions scenarios, and different future time periods. The potential for these model predictions to vary geographically presents a source of uncertainty in assigning climate-informed conservation strategies to landscapes. Regions with high agreement among predictions could be more confidently assigned a climate-informed strategy, whereas regions with less agreement among predictions may require a more cautious approach. More information on the data can be found at https://adaptwest.databasin.org/pages/uncertainty-climate-metrics.</p>
eConfidence School of Empathy game metrics
<p>Database includes metrics regarding in-game behaviour in eConfidence <em>School of Empathy </em>game. In supplementary file variable labels and descriptions are provided.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.