Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
88
datasets available to search
ShareScore release 0.9.0
Dataset results
88 results for “Domain Analysis”
Solar and interplanetary magnetic field data analyzed in "Optimal frequency-domain analysis for spacecraft time series: Introducing the missing-data multitaper power spectrum estimator"
<p>This dataset contains simultaneous measurements of the interplanetary magnetic field magnitude <B> and the sun's radio flux at 10.7 cm <F10.7>. <B> measurements come from a series of spacecraft located at the L1 point, while <F10.7> was measured by the ongoing monitoring program by Canada's Dominion Radio Astrophysical Observatory. Bartels rotation-averaged data were downloaded from NASA's OMNIWeb, https://omniweb.gsfc.nasa.gov/html/ow_data.html. The file contains other solar wind plasma parameters that were not used in the analysis.</p>
Analysis of EV Fires from Accesible Public Domain Data
<p>The incidents of EV fires have been collated from publicly accessible data, aiming to offer insights into potential EV fires in the future. The dataset also includes causes as determined from the reports analyzed.</p>
Dataset [Amazônia and Amazon: domain analysis with IRaMuTeQ in Scopus and LISA databases]
<p>The study reports the comparative analysis between the results of the search queries for the terms<em> Amazônia</em> and <em>Amazon</em> in Scopus and LISA databases, in the period from 2008 to 2018. Concept Theory and Domain Analysis were used in conjunction with IRaMuTeQ software in order to identify, quantify and analyse semantic distances in a sample consisting of 80 abstracts from retrieved articles.</p> <p>DOI: <a href="https://doi.org/10.5771/9783956507762-522">https://doi.org/10.5771/9783956507762-522</a></p> <p> </p>
Domain-Specific Language domain analysis and evaluation: a systematic literature review
<p>In order to successfully implement Domain-Specific Languages (DSLs), it is needed to systematically define and to support its development process; namely its Evaluation and the Domain Analysis phase. For that purpose, the studies were systematically selected from the most relevant venues that focus on the implementation of DSLs, in order to get insight if and how these development phases were performed. The special focus was given to the human-machine DSLs (excluding the machine-machine languages), the involvement of its end-users in the development process and the evaluation of DSLs usability, i.e. quality in use of DSLs. </p> <p>Preliminary results give us a notion that there is increased the state of practice of performing the evaluation of the DSLs, mostly including usability concerns, at least after its implementation. Generally, the quality of the reviewed studies was high. On another hand, rarely the assessments are done during domain analysis, which in general is not reporting inclusion of end-users or consideration of different use-cases, although the majority of studies refer to target non-programmers and to contribute easy in use. </p> <p>We did collect the valuable body of primary studies that are giving us answers to our research questions, however, to raise the credibility of the conclusions we should extend the analysis to other venues. Also, as we get insights into the categorization of practices we could specify more concretely answers that would give us means to perform more detailed meta-analysis. </p>
Unique composition and evolution histories of low velocity mantle domains: Data and analysis
<p>Dataset and Jupyter notebook accompanying 'Unique composition and evolution histories of low velocity mantle domains'. </p> <p>Dataset includes:</p> <ul> <li>Present day properties of simulated mantle for simulations RCY, B=0.22, B=0.44, visc2, visc3, CMB2600, CMB2800, COMP, PRM, MER.</li> <li>Present day predicted seismic properties for simulations RCY, B=0.22, B=0.44, visc2, visc3, CMB2600, CMB2800, COMP, MER.</li> <li>Present day predicted seismic properties for simulation PRM assuming 'primordial' material to be i) basaltic oceanic crust ii) chondrite enriched basalt (CEB).</li> <li>Present day delta Vs for simulations RCY, B=0.22, B=0.44, visc2, visc3, CMB2600, CMB2800, COMP, PRM, MER, filtered using the resolution of seismic tomography model S40RTS.</li> <li>Properties of simulated mantle at 100 Myr intervals from 900 Ma - 100 Ma inclusive for simulation RCY. </li> <li>P-T tables with predicted abundance of post-perovskite for different mantle lithologies (harzburgite, lherzolite and basalt - as defined in the paper).</li> </ul> <p>Jupyter notebook `s-llvps.ipynb` contains code for identifying simulated large low-velocity provinces (S-LLVPs), extracting assoicated model properties and plotting results. Python module files terra_utils.py and ppv.py are also included and required by the code in the notebook. </p> <p>There are a number of pre-requisite packages that will need to be installed in order to run the Jupyter notebook, including <a title="terratools" href="https://github.com/mantle-convection-constrained/terratools" target="_blank" rel="noopener">terratools</a>, a software package written specifically for reading and postprocessing outputs from TERRA simulations. Installation instructions can be found on the GitHub repository. </p> <p>Due to the TERRA code pre-dating open source licensing, we do not currently have permission to publicly share all aspects of the code. In code_pieces.F90 we include code snippets which were implemented for this study. </p> <p>Simulations were conducted using ARCHER2, the UK's national super-computing service. </p> <p>RCY.mp4 is a movie produced for simualtion RCY, visualising the evolution of temperature (right panels) and bulk composition (left panels). Hot iso-surface (red) drawn at +500 K and cold iso-surface (blue) drawn at -400 K, composition iso-surface drawn at C=0.6. Red and blue lines indicate overlying ridges / subduction zones taken from the plate motion reconstructions of Müller et al (2022). </p>
Prymnesium parvum 12B1 polyketide synthase (PKS) domain and module analysis
<p>The .zip file contains source data and code for an analysis that determines the modular structure from the domain structure of type I modular polyketide synthases (PKSs) in the Prymnesium parvum strain 12B1 gene annotation. In brief, a regex is used against a string representation of the PKS domain structures to determine the module structure. Some manual modifications for calling modules were performed where they deviated. Resulting data is available in the .zip file as .bed, .FASTA, .xlsx files. See Manuscript for further details.</p> <p>For versions after v1.0.5, analyses of B-type PKZILLA-B1 (i.e. from strain RCC3426) are also included. </p>
Domain Analysis Monitoring Public Dataset Dec 5
<p>Complex and heterogeneous software systems need to be monitored as their full behavior often only emerges at runtime, e.g., when interacting with other systems or the environment. Software monitoring approaches observe and check properties or quality attributes of software systems during operation. Such approaches have been developed in diverse communities and for various kinds of systems and purposes. For instance, requirements monitoring aims to check at runtime whether a software system adheres to its requirements, while resource or performance monitoring collects information about the consumption of computing resources by the monitored system. Many venues publish research on software monitoring, often using diverse terminology, focusing on different monitoring aspects and phases. The lack of a comprehensive overview of existing research often leads to re-inventing the wheel. We provide a domain model to structure and systematize the field of software monitoring, starting with requirements and resource monitoring. For more information on the domain model, please refer to the authors' publication lists. In this dataset we provide details on 47 approaches we analyzed with the model to assess its coverage.</p> <p>Please note that this is version 1.1 from January 30, 2019 and a more recent version might be available (link on top of dataset).</p>
Domain-Driven Design in Microservices-Based Systems Development: A Systematic Literature Review and Thematic Analysis [Dataset]
<p>This repository contains all artifacts related to the study: Domain-Driven Design in Microservices-Based Systems Development: A Systematic Literature Review and Thematic Analysis</p>
Annotation-based Modeling of Non-functional Requirements and Analysis Results in Domain-driven Design
<p>This repo contains all supplementary data sets that we have created and used throughout this thesis. In particular, it contains<br> - expert interview material (elicitation): consent form and question catalogue for requirements elicitation<br> - expert interview material (evaluation): consent form and task description for expert evaluation<br> - Diagrams related to our modeling concept and Dqualizer<br> - Screenshots of the Domain Story Modeler with our Modeling Concept</p>
Data from: Molecular evolutionary analysis of nematode Zona Pellucida (ZP) modules reveals disulfide-bond reshuffling and standalone ZP-C domains
<p>Zona pellucida (ZP) modules mediate extracellular protein-protein interactions and contribute to important biological processes including syngamy and cellular morphogenesis. While some biomedically-relevant ZP modules are well-studied, little is known about the protein family's broad-scale diversity and evolution. The increasing availability of sequenced genomes from "non-model" systems provides a valuable opportunity to address this issue, and to use comparative approaches to gain new insights into ZP module biology. Here, through phylogenetic and structural exploration of ZP module diversity across the nematode phylum, I report evidence that speaks to two important aspects of ZP module biology. First, I show that ZP-C domains—which in some modules act as regulators of ZP-N domain-mediated polymerization activity, and which have never before been found in isolation—can indeed be found as standalone domains. These standalone ZP-C domain proteins originated in independent (paralogous) lineages prior to the diversification of extant nematodes, after which they evolved under strong stabilizing selection, suggesting the presence of ZP-N domain-independent functionality. Second, I provide a much-needed phylogenetic perspective on disulfide bond variability, uncovering evidence for both convergent evolution and disulfide-bond reshuffling. This result has implications for our evolutionary understanding and classification of ZP module structural diversity and highlights the usefulness of phylogenetics and diverse sampling for protein structural biology. All told, these findings set the stage for broad-scale (cross-phyla) evolutionary analysis of ZP modules and position Caenorhabditis elegans and other nematodes as important experimental systems for exploring the evolution of ZP modules and their constituent domains.</p> <p> </p>
Phylogenetic analysis of Harmonin Homology Domains - Datasets
<p>Datasets associated to the article Phylogenetic analysis of Harmonin Homology Domains.</p> <p>HHD_starting-profile.hmm -> Profile HMM used to screen the UniprotKB</p> <p>HHD_all-hits_aligned.fa -> All hits aligned</p> <p>Other *.fa correspond to sequences identified for each cluster described in th article</p>
Molecular dynamics trajectories, GROMACS input files, and analysis code from "Rational optimization of a transcription factor activation domain inhibitor" by Basu et. al, Nature Structural & Molecular Biology, 2023
<p>Molecular dynamics trajectories, GROMACS input files, and analysis code from "Rational optimization of a transcription factor activation domain inhibitor" by Basu et. al, Nature Structural & Molecular Biology, 2023</p> <p> </p> <p> </p>
Binding of Cholesterol to the N-terminal Domain of the NPC1L1 Transporter: Analysis of the Epimerisation-Related Binding Selectivity and Loop Mutations
<p>Input files, topologies and trajectories of the work "Binding of Cholesterol to the N-terminal Domain of the NPC1L1 Transporter: Analysis of the Epimerisation-Related Binding Selectivity and Loop Mutations". </p>
Supplementary data: Effect of genotype by environment interaction (GEI) analysis for potato tuber yield and their quality traits in organic multi-environment domains of Poland
<p>Climate and raw data supplementary to the related publication in the journal Agriculture (ISSN 2077-0472).</p>
Data and data analysis codes for Palacios et al. "Single-domain Bose condensate magnetometer achieves energy resolution per bandwidth below ℏ"
<p>Data and data analysis codes for the article "Single-domain Bose condensate magnetometer achieves energy resolution per bandwidth below ℏ" by S. Palacios, et al. https://arxiv.org/abs/2108.11716</p>
The conductivity profile of Earth-ionosphere cavity used in the paper "Finite-difference time-domain analysis of ELF radio wave propagation in the spherical Earth-ionosphere waveguide and its validation based on analytical solutions" by Volodymyr Marchenko, Andrzej Kulak, Janusz Mlynarczyk
<p>The file "Marchenko_FDTD_Paper_Conductivity_Profile.dat" contains the conductivity profile of Earth-ionosphere cavity. The first column provides the altitude (in km) and the second column provides the conductivity (in S/m).</p>
Inputs and results of "A quantitative and qualitative citation analysis to retracted articles in the humanities domain"
<p>This repository contains the datasets and visualizations generated in our work: <strong>"A quantitative and qualitative citation analysis to retracted articles in the humanities domain"</strong>.</p> <p><strong>Note:</strong> the data are all contained inside the <strong><em>data.zip</em> </strong>file. You need to unzip the container to get access to all the files and directories listed below.</p> <p>The data (citations) gathered accompanied by their annotated characteristics are stored in <strong><em>data/</em>:</strong></p> <ul> <li><em>cits.csv: </em>a dataset containing all the entities (rows in the CSV) which have cited a retracted article in the humanities domain. Each citing entity (row) is accompanied by a set of features (columns) that characterizes it.<br> <strong>Note: </strong>this dataset is licensed under a <a href="https://creativecommons.org/publicdomain/zero/1.0/legalcode">Creative Commons public domain dedication (CC0)</a>.</li> <li><em>content.csv: </em>a dataset containing the abstracts and the in-text citation contexts of all the citing entities gathered.<br> <strong>Note: </strong>the data keep their original license (the one provided by their publisher). This dataset is provided in order to favor the reproducibility of the results obtained in our work.</li> <li><em>excluded_hum_retractions.csv: </em>a list of the 12 humanities retracted articles with a humanities affinity score < 2, therefore excluded from the analysis. </li> </ul> <p> </p> <p><strong>Topic modeling</strong></p> <p>We run a topic modeling analysis on the textual features gathered (i.e. abstracts and citation contexts). The results are stored inside the <em><strong>topic_model/</strong></em> directory. The topic modeling has been done using MITAO, a tool for mashing up automatic text analysis tools and creating a completely customizable visual workflow [1]. The directory <em><strong>workflow/ </strong></em>contains the workflows used in MITAO. The topic modeling results for each textual feature are separated into two different folders, <em><strong>abstract/</strong></em> for the abstracts, and <em><strong>cits_context/</strong></em> for the in-text citation contexts. Both the directories contain the following directories/files: </p> <ul> <li> <p><em><strong>datasets_and_views/: </strong></em>the datasets and visualizations generated using MITAO. </p> </li> <li> <p><em><strong>ldamodel_corpus_dict/: </strong></em>it contains the dictionary, the LDA topic model, and the tokenized and vectorized corpus.</p> </li> <li><em><strong>rawdata/: </strong></em>the textual collection, metadata, and stopwords used as input in the workflow of MITAO</li> </ul> <p> </p> <p><strong>References</strong></p> <p>[1] Ferri, P., Heibi, I., Pareschi, L., & Peroni, S. (2020). MITAO: A User Friendly and Modular Software for Topic Modelling [JD]. PuntOorg International Journal, 5(2), 135–149. <a href="https://doi.org/10.19245/25.05.pij.5.2.3">https://doi.org/10.19245/25.05.pij.5.2.3</a></p> <p> </p> <ol> </ol>
Annotated Dataset for Bilingual Code-Mixed English-Malay Sentiment Analysis and Sarcasm Detection in Public Security Domain
<p>Tweets from X, and post with comment from TikTok was acquired <span>from 11 September until 21 September 2022</span>. Data from both platforms was merged and selected. Three annotators manually labelling the selected data for sentiment and sarcasm. Sentiment labels are ‘positive’, ‘negative’, and ‘neutral’. Sarcasm label is ‘sarcastic’ and ‘not sarcastic’. Majority voting is considered for each label. Language identification label produced for each data. </p>
A narrative synthesis analysis of lifecourse domains within health service utilisation frameworks
<p>This paper is a scoping review which identifies existing frameworks, other than Anderson's Behavioural model, that look at Health Service Utilisation (HSU). In addition it investigates if these frameworks incorporate the use of a life course approach to understand HSU.</p> <p>The data included in this repository is:</p> <p>1. PRISMA flowchart</p> <p>2. PRISMA-ScR scoping review checklist</p> <p>3. Spreadsheet which contains all 70 frameworks (Table 6)</p> <p>4. References for the 70 Frameworks in Table6.</p> <p> </p> <p> </p>
An integrated analysis of the Passifloraceae virome using public-domain data
<p>This dataset is the result of an An integrated analysis of the Passifloraceae virome using public-domain data. </p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.