Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
2,373
datasets available to search
ShareScore release 0.9.0
Dataset results
2,373 results for “working”
Paleoclimate signals and groundwater age distributions from 39 public water works in the Netherlands; insights from noble gases and carbon, hydrogen and oxygen isotope tracers [Data set].
<p>Data set covering the meta data of the 39 well fields, the macro chemistry data and the data of the noble gases and carbon, hydrogen and oxygen isotope tracers used for assessing the paleoclimate signals and age distributions in the publication in Water Resources Research (2021)</p> <p><strong>Paleoclimate signals and groundwater age distributions from 39 public water works in the Netherlands; insights from noble gases and carbon, hydrogen and oxygen isotope tracers</strong></p> <p>Hans Peter Broers, Jürgen Sültenfuß<sup> </sup>, Werner Aeschbach, Arne Kersting,, Armin Menkovich, Jasperien de Weert and Jeroen Castelijns</p>
Appendices of the work "On the perceived relevance of critical internal quality attributes when evolving software features"
<p>Several refactorings performed while evolving software features aim to improve internal quality attributes like cohesion and complexity. Studies show that non-assisted refactorings might worsen, not improve, internal attributes. Current knowledge is scarce on how developers perceive the relevance of critical internal attributes while evolving features. Internal attributes are critical if their measurement assumes anomalous values. This qualitative study investigates the developer's perception on the relevance of critical internal attributes when evolving features. We target six class-level critical attributes: low cohesion, high complexity, high coupling, large hierarchy depth, large hierarchy breadth, and large size. We performed two industry case studies based on online focus group sessions. Developers discussed how much (and why) critical attributes are relevant for adding or enhancing features. We assessed the relevance of critical attributes individually and relatively, reasons behind the relevance of each critical attribute, and interrelations of critical attributes. Low cohesion and high complexity were perceived as very relevant because they often make evolving features hard while tracking failures and adding features. The other critical attributes were perceived as less relevant when reusing code or adopting design patterns. An example of perceived interrelation is high complexity leading to high coupling.</p>
Appendices of the work "On the perceived relevance of critical internal quality attributes when evolving software features"
<p>Several refactorings performed while evolving software features aim to improve internal quality attributes like cohesion and complexity. Studies show that non-assisted refactorings might worsen, not improve, internal attributes. Current knowledge is scarce on how developers perceive the relevance of critical internal attributes while evolving features. Internal attributes are critical if their measurement assumes anomalous values. This qualitative study investigates the developer's perception on the relevance of critical internal attributes when evolving features. We target six class-level critical attributes: low cohesion, high complexity, high coupling, large hierarchy depth, large hierarchy breadth, and large size. We performed two industry case studies based on online focus group sessions. Developers discussed how much (and why) critical attributes are relevant for adding or enhancing features. We assessed the relevance of critical attributes individually and relatively, reasons behind the relevance of each critical attribute, and interrelations of critical attributes. Low cohesion and high complexity were perceived as very relevant because they often make evolving features hard while tracking failures and adding features. The other critical attributes were perceived as less relevant when reusing code or adopting design patterns. An example of perceived interrelation is high complexity leading to high coupling.</p>
Data from: Capacity and selection in immersive visual working memory following naturalistic object disappearance
<p>Trial datasets and timeseries datasets associated with the experiment reported in the manuscript "Capacity and selection in immersive visual working memory following naturalistic object disappearance", by Babak Chawoush, Dejan Draschkow & Freek van Ede</p>
Datasets of the work named Development of a Low-Cost Smart Sensor GNSS System for Real-Time Positioning and Orientation for Floating Offshore Wind Platform
<pre>- 1_Motion_Simulator/ - IMU_results/ - 20211028101756.csv - 20220114101543.csv - 20220117000000.csv - Rotary_Table/ - 15/ - rover_20220105.nav - rover_20220105.obs - solution_20220105_CAS.log - 360/ - rover_20211221.nav - rover_20211221.obs - solution_20211221_CAS.log - 360-15/ - rover_20220202_CAS.nav - rover_20220202_CAS.obs - solution_20220202_CAS.log - Static_Tests/ - solution_SSRA00CAS0 - solution_SSRA00WHU0 - 2_GNSS_Signal_Simulator/ - platformmov_C1.xtd - platformmov_C2.xtd - platformmov_C3.xtd - TestBetaNoneMov_C1 - TestBetaNoneMov_C1.nav - TestBetaNoneMov_C1.obs - TestBetaNoneMov_C1.ubx - TestBetaNoneMov_C2 - TestBetaNoneMov_C2.nav - TestBetaNoneMov_C2.obs - TestBetaNoneMov_C2.ubx - TestBetaNoneMov_C3 - TestBetaNoneMov_C3.nav - TestBetaNoneMov_C3.obs - TestBetaNoneMov_C3.ubx - 3_Test_Sea/ - 20220503000000.xlsx - solution_28.nav - solution_28.obs - solution_28.ubx Background: {Journal Article using this dataset} 'Development of a Low-Cost Smart Sensor GNSS System for Real-Time Positioning and Orientation for Floating Offshore Wind Platform' Paper DOI: <a href="https://doi.org/10.3390/s23020925">https://doi.org/10.3390/s23020925</a> Abstract: a low-cost smart sensor GNSS system has been developed to provide accurate real-time position and orientation measurements on a floating offshore wind platform. The approach chosen to offer a viable and reliable solution for this application is based on the use of the well-known advantages of the GNSS system as the main driver for enhancing the accuracy of positioning. For this purpose, the data reported in this work are captured through a GNSS receiver operating over multiple frequency bands (L1, L2, L5) and combining signals from different constellations of navigation satellites (GPS, Galileo, and GLONASS), and they are processed through the precise point positioning (PPP) and real-time kinematic (RTK) techniques. Furthermore, aiming to improve global positioning, the processing unit fuses the results obtained with the data acquired through an inertial measurement unit (IMU), reaching final accuracy of a few centimeters. To validate the system designed and developed in this proposal, three different sets of tests were carried out in a (i) rotary table at the laboratory, (ii) GNSS simulator, and (iii) real conditions in an oceanic buoy at sea. The real-time positioning solution was compared to solutions obtained by post-processing techniques in these three scenarios and similar results were satisfactorily achieved. </pre>
Data of Chinese treatment group for the research work "Disentangling material, social, and cognitive determinants of human behavior and belief".
<p>This repository contains data files of Chinese treatment group for the research work "Disentangling material, social, and cognitive determinants of human behavior and belief".</p>
RDF dataset produced in the work "Exploring Adverse Outcome Pathways for Nanomaterials with semantic web technologies"
<p>Adverse Outcome Pathways (AOPs) have been proposed to facilitate mechanistic understanding of interactions of chemicals/materials with biological systems. Each AOP starts with a molecular initiating event (MIE) and possibly ends with adverse outcome(s) (AOs) via a series of key events (KEs). So far, the interaction of engineered nanomaterials (ENMs) with biomolecules, biomembranes, cells, and biological structures, in general, is not yet fully elucidated. There is also a huge lack of information on which AOPs are ENMs-relevant or -specific, despite numerous published data on toxicological endpoints they trigger, such as oxidative stress and inflammation. We propose to integrate related data and knowledge recently collected. Our approach combines the annotation of nanomaterials and their MIEs with ontology annotation to demonstrate how we can then query AOPs and biological pathway information for these materials. We conclude that a FAIR (Findable, Accessible, Interoperable, Reusable) representation of the ENM-MIE knowledge simplifies integration with other knowledge.</p>
Example data for working with the ASpecD framework
<p>ASpecD is a Python framework for handling spectroscopic data focussing on reproducibility. In short: Each and every processing step applied to your data will be recorded and can be traced back. Additionally, for each representation of your data (e.g., figures, tables) you can easily follow how the data shown have been processed and where they originate from.</p> <p>To provide readers of the publication describing the ASpecD framework with a concrete example of data analysis making use of recipe-driven data analysis, this repository contains both, a recipe as well as the data that are analysed, as shown in the publication describing the ASpecD framework:</p> <ul> <li>Jara Popp, Till Biskup: ASpecD: A Modular Framework for the Analysis of Spectroscopic Data Focussing on Reproducibility and Good Scientific Practice. Chemistry--Methods 2:e202100097, 2022. doi:10.1002/cmtd.202100097</li> </ul>
Difference and number of works published over the years grouped by the objective of the generative process
<p>Difference and number of works published over the years grouped by the objective of the generative process. Part of the study "What do we mean by GenAI?"</p>
Number of works grouped by AI technique employed
<p>Number of works grouped by AI technique employed. Part of the study "What do we mean by GenAI?"</p>
Number of works grouped by objective, domain and AI technique employed
<p>Number of works grouped by objective, domain and AI technique employed. Part of the study "What do we mean by GenAI?"</p>
Number of works published over the last five years
<p>Number of works published over the last five years. Part of the study "What do we mean by GenAI?"</p>
Number of works grouped by domain of application
<p>Number of works grouped by domain of application. Part of the study "What do we mean by GenAI?"</p>
Relationships among the generated content type, task, AI technique, and application domain in the retrieved works
<p>Relationships among the generated content type, task, AI technique, and application domain in the retrieved works. Part of the study "What do we mean by GenAI?"</p>
Number of works published over the years grouped by generated content type
<p>Number of works published over the years grouped by generated content type. Part of the study "What do we mean by GenAI?"</p>
PALaC Working Data
<p>PALaC's woking data, consisting in the lists, transcriptions and annotated materials we used for the project. Furthermore, we also include all the maps developed for the historical worck package of the project.</p> <p>Please make sure you read the READ ME FIRST file.</p>
Artifacts supplementing the ACM DTRAP 2020 article "Will You Trust This TLS Certificate? Perceptions of People Working in IT (extended version)"
<p>These research artifacts supplement the following two publications:</p> <ul> <li>Will You Trust This TLS Certificate? Perceptions of People Working in IT [ACSAC 2019], DOI 10.1145/3359789.3359800, more details at https://crocs.fi.muni.cz/public/papers/acsac2019</li> <li>Will You Trust This TLS Certificate? Perceptions of People Working in IT (extended version) [ACM DTRAP 2020], DOI 10.1145/3419472, more details at https://crocs.fi.muni.cz/public/papers/dtrap2020</li> </ul> <p>The artifacts contain the full experimental setup (as described in Section 2.1 of the paper) and the complete anonymized dataset underlying the evaluation presented in Sections 3 and 4.</p> <p>The experimental setup contains the documents accompanying the task: the informed consent, pre-task questionnaire, task description, trust scales, and the list of questions posed during the post-task interview (all in PDFs). We further include the custom website with certificate validation documentation for the “redesigned” condition (static HTML). While working on the task, participants in the “redesigned” condition could access this website via a link that was in the redesigned error messages. Furthermore, we provide the software with which the participants interacted. It contains the displayed error messages and validated certificates. These things are available both individually and incorporated in a snapshot of a virtual machine used at the experiment (importable directly into VirtualBox).</p> <p>The collected data is presented in a single dataset (SPSS format; you can use PSPP as a free alternative). It includes the analysis syntax files to obtain the numerical results presented in the paper. For each participant, the dataset contains: 1) pre-task questionnaire answers, 2) reported trust ratings, 3) sub-task timing, 4) information on whether they browsed the Internet and 5) the interview codes assigned. Note that we do not publish the interview transcripts to preserve participant privacy.</p>
Artifacts supplementing the RSA-CT 2018 paper "Why Johnny the Developer Can't Work with Public Key Certificates"
<p>Supplemental materials for the paper "Why Johnny the Developer Can't Work with Public Key Certificates" (DOI 10.1007/978-3-319-76953-0_3, more details at https://crocs.fi.muni.cz/public/papers/rsa2018) contain the following:</p> <ul> <li>Informed consent participants had to sign (experiment design approved by Research Ethics Committee of Masaryk University)</li> <li>General questionnaire & System usability scale questionnaire</li> <li>User tasks & certificates to validate</li> </ul>
TrainTicket microservice testbench extracted information for our work: Evaluating ChatGPT's Proficiency in Understanding and Answering Microservice Architecture Queries Using Source Code Insights
<p>It contains the CSV file output of our tool implemented in the paper: "Evaluating ChatGPT’s Proficiency in Understanding and Answering Microservice Architecture Queries Using Source Code Insights." applied to the TrainTicket microservice testbench. The information in this CSV was used for In-Context-Learning for ChatGPT.</p>
HeatResilientCity II - work package 2.2: Influence of regional and urban climate on indoor overheating - Results of building performance simulation
<p>This repository contains the <strong>results of the building performance simulations</strong> carried out in the working package 2.2 Influence of regional and urban climate on indoor overheating of the project <a href="http://heatresilientcity.de/">HeatResilientCity II</a>. The buildings under consideration are a multi-residential so-called ‘Gründerzeithaus’ (GZH) and a large-panel construction (LPC) building. The results were extracted for two rooms on the top floor/attic of each building and follow a consistent name convention. Each file contains hourly resolved values for the outdoor air temperature, the indoor air temperature, the indoor operative temperature and the relative humidity indoors and outdoors. Further information can be found in the README of this repository. The simulations were performed for five <strong>different</strong> <strong>regions</strong> in Germany (Dresden, Hamburg, Köln, Stuttgart and Potsdam) for <strong>average present</strong> and <strong>future</strong> <strong>summers</strong> based on meteorological measurement data and under consideration of <strong>urban</strong> <strong>climate</strong>.</p> <p>In addition to the ‘plain’ simulation results, some <strong>heat-indicator variables</strong> were calculated and listed in the files <em>Calculated_Variables.txt</em>. The calculated quantities include temperature-weighted exceedance hours (TWEH) for the limits of 25, 26 and 27 °C (defined in DIN 4108-2:2013 as ‘Übertemperaturgradstunden’) and the maximum operative temperature calculated for the period from April to September.</p> <p>The used <strong>input data</strong> and <strong>building models</strong> can be found in the related repository.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.