Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
118
datasets available to search
ShareScore release 0.9.0
Dataset results
118 results for “Software Modelling”
P-S waves 3D velocity model of Los Humeros area from earthquake based travel-time tomography using CAT3D software (OGS)
<p>The dataset contains the 3D velocity model (VP (m/s), VS (m/s) and VP/VS) obtained from the tomographic inversion of seismological data in the area of Los Humeros (Mexico). The model was performed in the frame of the GEMex project (Mexico‐Europe Cooperation for research of enhanced geothermal systems and super-hot geothermal systems, WP5 ‘Detection of deep structures’, Jousset et al., D5.3, 2019).</p> <p>The inversion used 2661 P arrivals and 2272 S arrivals associated to 395 earthquakes recorded by 37 stations. The picking data was provided by Toledo et al., 2019.</p> <p>The inversion was performed by CAT3D software, a tomographic tool developed by OGS, which uses the SIRT method (Simultaneous Iterative Reconstruction Technique, Stewart, 1993) as inversion algorithm and the ray tracing procedure based on minimum time principle (Böhm et al., 1999). The velocities used as initial model for tomography were provided by the interpolated values obtained from the velocity analysis of four 2D seismic lines acquired inside the same investigated area by the tomographic inversion (See GEMex deliverable D5.3).</p> <p>The 3D velocity model is defined by a 3D grid of 61 nodes in X, 69 nodes in Y and 29 nodes in Z, equally spaced by 250 m in all directions. The total dimensions of the model is 15x17x7 km and the borders positions are (m) (WGS 84/UTM ZONE 14N):</p> <p>Xmin = 655000, Xmax = 670000</p> <p>Ymin = 2168000, Ymax = 2185000</p> <p>Zmin = -3000, Zmax = 4000</p>
Dataset for : A New Era in Software Security: Towards Self-Healing Software via Large Language Models and Formal Verification
<p>We present a novel solution combining Large Language Model (LLM) capabilities with Formal Verification strategies to falsify and automatically repair software vulnerabilities. Initially, we employ Bounded Model Checking (BMC) to locate the software vulnerability and derive a counterexample. Relying on mathematical proofs, counterexamples provide evidence that the system behaves incorrectly or contains a vulnerability, thereby preventing the generation of false positive alerts. The counterexample that has been detected, along with the source code, are provided to the LLM engine. Our approach involves establishing a specialized prompt language for conducting code debugging and generation to understand the vulnerability's root cause and repair the code. Finally, we use BMC to verify the corrected version of the code generated by the LLM. As a proof of concept, we create \esbmcai based on the Efficient SMT-based Context-Bounded Model Checker (ESBMC) and a pre-trained Transformer model, specifically gpt-3.5-turbo, to detect and fix errors in C programs. We generated a dataset comprising $1{,}000$ C code samples, each consisting of $20$ to $50$ lines of C code. Experimental results show that our proposed method achieved an impressive success rate of up to $80$\% in repairing vulnerable code, encompassing buffer overflow, arithmetic overflow, and pointer dereference failures. To our knowledge, \esbmcai represents the first proposal for a pioneering initiative to integrate a Large Language Model (LLM) with software model checking. We advocate that this automated approach has the potential to incorporate into the software development lifecycle's continuous integration and deployment (CI/CD) process. </p> <p> </p> <p>The uploaded dataset contains 1000 codes, each comprising 20 to 50 lines of C code generated with gpt-3.5-turbo. The material also consists of a version of ESBMC statically compiled with all dependencies, a classifier script, and the output file.</p> <p> </p> <p> </p>
Testing 3D modelling software. Modelling charging pads for WPT of electric vehicles for EM emissions simulation.
<p>Even for the experienced 3D FEM modelers it may not be obvious which geometry discretization is the most appropriate and suitable for this type of problem. It may be a conservative approach to test the computation tool on a simplified geometry, on which the magnetic field distribution is known. As part of the “Metrology for inductive charging of electric vehicles” (MICEV) project (www.micev.eu), an axisymmetric geometry was used, with the results reported.</p>
Survey on the usage of Mathematical Modelling, Simulation and Optimization software
<p>This dataset contains the result of a survey we carried out in the context of the MSO4SC project in order to know which kinds of tools for simulation were using our stakeholders. The purpose was to prioritize functionalities depending on stakeholders' preferences. It was a survey with 41 questions grouped in 10 areas (impact of simulation software on their entities, usage of pre/post-processing, usage of visualization, etc...). The pdf file includes the list of questions for clarification. Such survey was answered by academia and industry from several European countries.</p>
Estimating heavy metal deposition in Germany using model calculations and biomonitoring data, link to research data and scientific software
<p>Research data and scientific software related to an investigation dealing with modelled data on Cd and Pb deposition (LOTOS-EUROS, EMEP/MSC-East) and monitoring data from the International Cooperative Programme on Effects of Air Pollution on Natural Vegetation and Crops (ICP Vegetation Moss Survey) and the German Environmental Specimen Bank (ESB) providing corresponding parameters on HM concentration in various biota. The study aimed at examining, whether an integrated use of model calculations and monitoring data can extend established methods for estimating and evaluating spatial patterns of atmospheric Pb and Cd deposition across Germany.</p>
Fuzzy modelling and mapping soil moisture in Germany, link to research data and scientific software
<p>Research data and scientific software related to spatio-temporal estimations of ecological soil moisture with available data covering the whole territory of Germany and the Kellerwald National Park (Hesse). Temporal trends of modelled soil moisture for the time period 1961–2070 were statistically analyzed. Soil moisture changes (drying-out) at both national and regional levels were mapped.</p>
A Modified Doyle-Fuller-Newman Model Enables the Macroscale Physical Simulation of Dual-ion Batteries - Dataset and Software
<p>This dataset contains:</p> <p>- all the raw cycling data of the three-electrode cell used to gather the experimental data for the model validation (VMP data, exported with EC-LAB);<br>- the specific, processed data used in the model validation step (0.2C discharge, 5C discharge, EIS data);<br>- the COMSOL dual-ion battery model (version 6.0). IMPORTANT: activate the "Electric potential at the positive electrode current collector (only for EIS)" boundary condition when simulating impedance spectroscopy, and deactivate it when simulating charge/discharge curves; the charge-discharge profile can be modified by changing the duration of the test, the C-rate, and the conditions set in the "Events" section.</p> <p>Update: Fixed the model to work also in the 6.2 version of COMSOL (Substituted Dleff with Dleffxx in the modified weak expression of the cathode mass conservation equation). Download the new version!</p>
Data, code and software to reproduce the article entitled "Modeling soil-plant functioning of intercrops using comprehensive and generic formalisms implemented in the STICS model"
<p>This is the data, code and software to reproduce the article entitled " Modeling soil-plant functioning of intercrops using comprehensive and generic formalisms implemented in the STICS model". Here is a summary of the paper:</p> <p>The growing demand for sustainable agriculture is raising interest in intercropping for its multiple potential benefits to avoid or limit the use of chemical inputs or increase the production per surface unit. Predicting the existence and magnitude of those benefits remains a challenge given the numerous interactions between interspecific plant-plant relationships, their environment and the agricultural practices. Soil-crop models are critical in understanding these interactions in dynamics during the whole growing season, but few models are capable of accurately simulating intercropping systems.</p> <p>In this study, we propose a set of simple and generic formalisms for simulating key interactions in intercropping systems that can be readily included into existing dynamic crop models. This requires simulating important processes such as development, light interception, plant growth, N and water balance, and yield formation in response to management practices, soil conditions, and climate. These formalisms were integrated into the STICS soil-crop model and evaluated using observed data of intercropping systems of cereal and legumes mixtures, including Faba bean-Wheat, Pea-Barley, Sunflower-Soybean, and Wheat-Pea mixtures. We demonstrate that the proposed formalisms provide a comprehensive simulation of soil-plant interactions in various types of bispecific intercrops. The model was found consistent and generic under a range of spring and winter intercrops (nRMSE = 25% for maximum leaf area index, 23% for shoot biomass at harvest, and 18% for yield).</p> <p>This is the first time a complete set of formalisms has been developed and published for simulating intercropping systems and integrated into a soil-crop model. With its emphasis on being generic, sufficiently accurate, simple, and easy to parameterize, STICS is well-suited to help researchers designing <em>in silico</em> the agroecological transition by virtually pre-screening sustainable, manageable intercrop systems adapted to local conditions.</p> <p> </p> <p> </p> <p> </p>
Simulation data and software scripts used in calculus of ∆36 signature from EMAC clumped O2 isotope-inclusive model
<p>This publication contains simulation data and software scripts for calculating quantities related to clumped oxygen isotope signature (∆<sub>36</sub>) derivation, as described in the "static" framework of Yeung‍ et‍ al. (2016), hereinafter "Y16") and subsequently used in Yeung‍ et‍ al.‍ (2019) analysis. We provide the output of the 1950–2011 transient simulation with EMAC model with explicit "dynamic" simulation of ∆<sub>36</sub> (i.e. <sup>18</sup>O<sup>18</sup>O isotopologues undergoing transport, mixing and O(<sup>3</sup>P)-mediated isotope equilibration) to demonstrate the importance of several assumptions/simplifications involved in the static calculus.</p> <p> </p> <p>Please refer to .README.pdf for details.</p>
Dataset and Analysis Scripts for Survey "Understanding Security Tactics in Microservice APIs using Annotated Software Architecture Decomposition Models -- A Controlled Experiment"
<pre>Dataset, R-Scripts and questionnaire templates for our survey <em>Understanding Security Tactics in Microservice APIs using Annotated Software Architecture Decomposition Models -- A Controlled Experiment.</em></pre>
Software and data underlying the article 'A serious game approach for lake modeling and management: the EscapeBLOOM'
<p>Here we share the player version of the EscapeBLOOM, a dummy version showcasing the techniques to create a similar digital escape room, and the anonymized data of the quantitative survey as presented in the publication 'A serious game approach for lake modeling and management: the EscapeBLOOM'.</p> <p>Anyone is free to play or adjust the game for their own educational purposes. The dummy and supplementary material of the publication 'A serious game approach for lake modeling and management: The EscapeBLOOM' <a title="Persistent link using digital object identifier" href="https://doi.org/10.1016/j.envsoft.2024.105941" target="_blank" rel="noreferrer noopener">https://doi.org/10.1016/j.envsoft.2024.105941</a> together provide guides on how to create a new game from the start and may help to adjust the existing game.</p> <p>The data of the survey was used for the analysis of perceived learning in the publication 'A serious game approach for lake modeling and management: the EscapeBLOOM'.</p>
JSON files containing parameters of training gene models for ab-initio prediction software
<p>These are the JSON files containing parameters of training gene models for ab-initio prediction software. These training datasets are Phytophthora specific and can be further utilized for the gene prediction and annotation of other related Phytophthora strains.</p>
Dataset: We Do Not Understand What It Says -- Studying Student Perceptions of Software Modelling
<p>The dataset contains two supporting documents for the paper title, "We Do Not Understand What It Says -- Studying Student Perceptions of Software Modelling". The first one is an excel sheet containing interview transcripts of 13 of the participants of this study (who agreed to publish their statements) and the second is an appendix file containing the interview guide (questionnaire used for interviews with students and instructors) used in the case study. </p> <p>The interview transcripts are supported by "in-vivo coding" used by both authors separately during analysis. </p>
On the Application of Machine Learning Models to Assess and Predict Software Reusability
<p>This is the dataset, results and notebook for the submission into Maltesque 2022 conference.</p>
Dataset and Replication Package for the View-Based Retriever Approach To Reverse Engineering Software Architecture Models
<div> <div><span>Dataset and replication package for the view-based Retriever approach to reverse engineering software architecture models. Each Dataset project is structured as follows:</span></div> <ul> <li><span>The .ruleengine.yml file contains the configuration for running the Retriever approach.</span> <ul> <li><span>The repository value is the ID of a GitHub repository.</span></li> <li><span>The current_version value is the latest version of the retriever approach used to build the architectural models.</span></li> <li><span>The rules values are the rules used to build the architectural models.</span></li> </ul> </li> <li><span>The model_re folder contains the architectural model of the system automatically generated by the Retriever approach.</span> <ul> <li><span>The pcm folder contains the Palladio Component Model (PCM) of the system.</span></li> <li><span>The uml folder contains the PlantUML model.</span></li> </ul> </li> <li><span>The model_gs folder contains our manual gold standards for the system.</span></li> </ul> <div><span>The easiest way to use our approach is to use the CLI application with the given parameters: ./eclipse -i /path/to/input/directory -o /path/to/output/directory -r supported_rules</span></div> </div>
PREPCLIM software - additional material for publication in "Geoscientific Model Development" 2024
<p>Data set illustrating the software developed in the PREPCLIM project and used in the proposed paper:</p> <p>A Modeling System for Identification of Maize Ideotypes, optimal sowing dates and nitrogen<br>fertilization under climate change – PREPCLIM-v1</p> <p>https://doi.org/10.5194/gmd-2024-105<br>Preprint. Discussion started: 11 July 2024<br>c Author(s) 2024. CC BY 4.0 License.</p>
Figure 1. Markov Chain Model&Figure 2. Transition matrix-Study of a Random Navigation on the Web Using Software Simulation
<p>For a good simulation it is very important to find methods for<br> navigating through the web (Levene and Wheeldon, 2004). John Kemeny and Laurie Snell have<br> proposed the use of Markov models for web simulations (Kemeny and Snell, 1960). Cadez et al. (2000)<br> used Markov models for classifying the sessions into different categories for browsers. Some other<br> proposed techniques choose to combine different order Markov models for obtaining low state<br> complexity and improving accuracy, as Deshpande and Karypis (2004). Dongshan and Junyi (2002)<br> used for predicting the access providing good scalability and high coverage a hybrid-order tree-like<br> Markov model. As an alternative to the Markov model Pitkow proposed a longest subsequence model<br> (Pitkow and Pirolli, 1999), also for predicting the next page accessed by the user Sarukkai chose<br> Markov models (Sarukkai, 2000).<br> Transitions are simulated using the Markov Chain nodes, Google matrix and an arbitrary initial<br> probability distribution. Examples can be seen in Figure 1 and Figure 2.</p>
The Fundamentals Regarding the Usage of the Concept of Interface for the Modeling of the Software Artefacts-Figure 4. The artefact during the initial structuring stage
<p>The IT developer elaborates the detailed structure of the artefact, while considering several aspects:</p> <p>— The resources that are necessary to the artefact in order to accomplish its mission;</p> <p>— The artefact’s resistance to the changes regarding the functional requirements;</p> <p>— The artefact’s resistance to the technological changes;</p> <p>— The reasonably priced integration of the artefact in the structure of the host system;</p> <p>— The assurance of a reasonable reusability coefficient of the artefact during the struc- turing process of other artefacts;</p> <p>— The flexibility of the relations that exist among the components of the artefact;</p> <p>— The flexibility of the artefact’s connections with the host system.</p> <p> </p>
Figure 3. The artefact as it exists as a black box-The Fundamentals Regarding the Usage of the Concept of Interface for the Modeling of the Software Artefacts
<p>Consequently, the artefact as it exists as a black box can be represented according to the representation in Figure 3. It can be noticed that the artefact as it exists as a black box begins to interact with the environment. Two main categories of interfaces may be utilized by any artefact in order to interact with the environment: — Human Computer Interfaces (HCI); — Shared Resource Interfaces (SRI).</p>
Figure 1. Visual and synthetic representation of the modelling process in the software industry-The Fundamentals Regarding the Usage of the Concept of Interface for the Modeling of the Software Artefacts
<p>The experience that is accumulated regarding the modelling paradigms in the software engineering is impressive. Thus, the software engineering recognizes modelling paradigms like object orientation, aspect orientation, component orientation, service orientation, agent orientation. In one form or another, these paradigms prove their ex- cellence in certain types of IT projects. At the same time, these paradigms reveal their objective limits when they are used to engineer the real world software systems. Every modelling paradigm represents, in fact, a modality to represent the real world using a specific formal framework. The specificity of the formal framework is defined from both a syntactic and semantic perspective. The formal syntactic framework of a paradigm refers to the concepts that are used by the paradigm in order to represent the real world, but also to the recommended principles that allow for these concepts to interact in a correct and efficient manner. Both the concepts and the principles benefit from a formal representation that ultimately favours communication as a secondary modelling lever inside the IT projects. Every syntactic artefact of a paradigm can be associated with a certain real world semantics, which it abstracts. As a consequence, considering that the real world continuously enhances its semantic potential, the syntactic constructs that are favoured by the paradigm may become problematic.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.