Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
4,486
datasets available to search
ShareScore release 0.7.1
Dataset results
4,486 results for “exploration”
The commitment to global sea level rise over the next 500 years: exploring the threat of the Antarctic Ice Sheet to coastal infrastructure
<p>Within Australia alone, more than A$226 billion of coastal infrastructure is vulnerable to the anticipated rise in sea level by the end of the century. The IPCC Fifth Assessment Report concludes that the likely increase in global mean sea level during the 21st century ranges from 26-55 centimetres (under the low-end RCP2.6 climate scenario) to 45-82 centimetres (under the high-end RCP8.5 climate scenario). However, these projections do not take into account the potential for collapse of the marine-based sectors of the Antarctic Ice Sheet.</p> <p>Recent evidence has indicated that the IPCC projections may be under-estimates, with sea level increases of up to 2.5 metres possible by the end of the 21st century. Modelling studies have also demonstrated the potential for the Antarctic Ice Sheet to undergo irreversible collapse during the coming centuries, leading to dramatic increases in global sea level on time scales relevant to critical coastal infrastructure such as refineries and airports. The most extreme prediction is that Antarctica could contribute 15.65±2.00 metres to global sea level by the year 2500.</p> <p>Here, we combine climate modelling and ice sheet modelling to explore the evolution of the Antarctic Ice Sheet over the next 500 years under a range of climate scenarios. We run the models many times to take into account gaps in our understanding of ice sheet dynamics. This allows us to generate robust projections of the Antarctic contribution to global sea level from the present to the year 2500, complete with quantified confidence intervals. We conclude that the sea level contribution during the 21st century will be modest, consistent with the IPCC Fifth Assessment Report, but that melting of the Antarctic Ice Sheet will accelerate thereafter. By the year 2500, we predict that the Antarctic contribution to global sea level will be at least 5 metres.</p>
Audio clips of Orca (Orcinus orca) and non-orca sounds for the exploration of multiple acoustic representations
<p>Data and code associated with "Comparing acoustic representations for deep learning-based classification of underwater acoustic signals: a case study on orca (Orcinus orca) vocalizations."</p> <p>A collection of 9600 audio clips recorded by a hydrophone off San Juan Island, WA, USA. The clips are 3 seconds in duration with a sampling rate of 64KHz, and contain a variety of orca vocalizations (in the srkw folder), as well as non-orca sounds, both humpbacks (hb folder) and unspecified sounds typical of the location (neg folder). </p> <p>The code for each of the representations used in this study is also included.</p>
Exploring the critical zone heterogeneity and the hydrological diversity using an integrated ecohydrological model in three contrasted long-term observatories
<p>These files provide useful data and supplementary material associated with the publication 'Exploring the critical zone heterogeneity and the hydrological diversity using an integrated ecohydrological model in three contrasted long-term observatories' (MNT information, atmospheric forcings, R scripts used to process and draw the graphs from the EcH2O-iso simulations, and observed water discharges).</p>
Dataset for manuscript titled "Exploring SureChEMBL from a drug discovery perspective".
<p>This is the data directory for running the code available on the GitHub repository for the manuscript titled "Exploring SureChEMBL from a drug discovery perspective<strong></strong>". The GitHub repository is available at <a href="https://github.com/Fraunhofer-ITMP/patent-clinical-candidate-characteristics">https://github.com/Fraunhofer-ITMP/patent-clinical-candidate-characteristics.</a></p>
Explore maps as you read a comic book
<p>This repository contains three documents accompanying the article 'Explore maps as you read a comic book.</p> <p>The first document compiles several types of breaks encountered in various pan-scalar maps. Each type is also assigned a progression note after an analysis of their effects on the user.</p> <p>The next table illustrates an example of good generalisation for each of the hydrographic patterns observed in the article. This generalisation is based on three key concepts of progressivity, promoting a continuity of meaning, coherence, and rhythm.</p> <p>The final document represents a scale master inspired by the methodology of (Brewer and Buttenfield, 2007). In the article, we demonstrate how to adapt it to pan-scalar maps to better visualize sequences, as well as two types of rhythms: map cadence and abstraction cadence.</p>
Exploring localized ENZ resonances and their role in superscattering, wideband invisibility, and tunable scattering
<p>The files contain the data generated by MATLAB and a sample MATLAB code for the selected figures. </p> <p>Research founded by Narodowe Centrum Nauki, project no UMO-2020/39/I/ST3/02413.</p>
Raw data for PIP2 interaction with TRPC3, explored through computation and electrophysiology
<p>The transient receptor potential canonical type 3 (TRPC3) channel plays a pivotal role in regulating neuronal excitability within the brain via its constitutive activity. The channel is intricately regulated by lipids and has previously been demonstrated to be positively modulated by PIP2. Using molecular dynamics simulations and patch clamp techniques, we reveal that PIP2 predominantly interacts with TRPC3 at the L3 lipid binding site, located at the intersection of pre-S1 and S1 helices. We propose a novel signal transduction pathway from the L3 through the re-entrant loop to a salt bridge between the TRP helix and S4-S5 linker. Notably, we find that both stimulated and constitutive TRPC3 activity require PIP2. These structural insights into the function of TRPC3 are invaluable for understanding the role of the TRPC subfamily in health and disease in native tissue.</p>
Dataset for exploring design patterns in Qiskit programs
<p>This set of plots belongs to the magazine article about a exploration of design patterns in quantum software done by Miriam Fernández-Osuna, Ricardo Pérez-Castillo, José Antonio Cruz-Lemus, Michal Baczyk and Mario Piattini, Grupo Alarcos Research Group.</p>
Exploring the Exclusive Isolation of Pseudomonas syringae in Peltigera Lichens via metabolite analysis and growth assays - Appendix
<p>Lichen samples from Iceland were collected from the genera Peltigera, Cladonia, and Stereocaulon in March 2023 at Heidmork forest, Oskjuhlid hill, and the shores of Ellidaa in Arbaejarstifla. All specimens underwent morphological analysis, and corresponding vouchers have been deposited at the Icelandic Institute of Natural History.</p>
Supporting Data for "Exploring ChatGPT-4 for Transforming Taxonomic Data into OWL: Lessons Learned and Implications for Ontology Development"
<p>Data from the trials with ChatGPT to generate OWL files for taxonomic data from the GBIF Backbone Taxonomy.</p> <p>Updates of version 2: additional prompts from the experiments with Gemini and DeepSeek.</p>
Exploring the Pocillopora cryptic diversity: a new genetic lineage in the western Indian Ocean or remnants from an ancient one?
<p>Cryptic species and lineages have been widely reported during the last decades, particularly in the marine realm. Misidentifications and ignoring species complexes imply many consequences, notably biasing biodiversity and connectivity assessments, which in turn mislead our understanding of ecosystems and impact the effective design and management of conservation plans. Focusing on the Indo-Pacific coral genus <em>Pocillopora</em>, playing key roles in reef ecosystems as one of the main bio-constructors, we report the first <em>Pocillopora</em> PSH16 (ORF53; <em>sensu</em> Gélin et al. 2017, Mol Phylogenet Evol 109:430–446) colonies (<em>N</em> = 19) in the western Indian Ocean (Nosy Tanikely, Madagascar), 6,000 km further from its current distribution. Colonies were identified according to their mitochondrial Open Reading Frame (ORF) haplotype and Bayesian assignment tests based on 13-microsatellite genotypes. Additionally, we performed genetic structure and diversity analyses with sympatric colonies from other <em>Pocillopora</em> species and <em>Pocillopora</em> PSH16 colonies from the tropical southwestern Pacific, revealing (1) a weak clonal richness, (2) a weak genetic diversity and (3) a relative isolation for the newly reported PSH16 colonies. These colonies thus represent either a new, distinct and uncommon, genetic lineage, or isolated remnants of a wider one. In any case, unless specific management measures are implemented, their long-term maintenance seems compromised due to restricted gene flow within a restricted pool of genes.</p> <p> </p> <p>This dataset contains the microsatellite genotypes analysed (98 <em>Pocillopora</em> colonies × 13 loci + ORF). Missing data are encoded as "?". The sampling marine province and the population are indicated for each individual.</p>
Inputlog Copy Task Corpus: Exploring and defining typing skills
<p><strong>Context</strong></p> <p>One of the components that is included in the keystroke logging program Inputlog (<a href="https://www.inputlog.net">https://www.inputlog.net</a>) is the Copy Task component. It consists of a multi-layered set of tasks that measure a person's typing skill:</p> <table> <tbody> <tr> <td>Tapping task</td> <td>press the ‘d’ and ‘k’ key alternatively during 15 s</td> </tr> <tr> <td>Sentence</td> <td>copy a sentence during 30 s</td> </tr> <tr> <td>Word combination 1</td> <td>copy a combination of three words seven times</td> </tr> <tr> <td>Word combination 2</td> <td>copy a combination of three words seven times</td> </tr> <tr> <td>Word combination 3</td> <td>copy a combination of three words seven times</td> </tr> <tr> <td>Word combination 4</td> <td>copy a combination of three words seven times</td> </tr> <tr> <td>Consonant groups</td> <td>copy four blocks of six consonants once</td> </tr> </tbody> </table> <p>The task is currently made available in twelve languages. </p> <p>For more information: <a href="https://doi.org/10.5334/jors.234 ">https://doi.org/10.5334/jors.234 </a></p> <p> </p> <p><strong>Interactive Dashboard</strong><br> Visit the webpage with an interactive dashboard to explore, filter, and download the +5K copy task corpus.</p> <p><em><strong>website</strong></em>: <a href="https://www.inputlog.net/copy-task/">https://www.inputlog.net/copy-task/</a><br> <em><strong>dashboard</strong></em>: <a href="https://inputlog-analysis.uantwerpen.be/expert">https://inputlog-analysis.uantwerpen.be/expert</a></p> <p> </p> <p><strong>Corpus</strong></p> <p>We are happy to make a multilingual corpus available (open access) that currently consists of more than 5000 copy tasks. </p> <ul> <li>The + 5K corpus is carefully cleaned and fully anonymized.</li> <li>The Shiny interface allows users to filter the corpus based on about 10 variables.</li> <li>The selection can be downloaded in different formats and levels of aggregation (from raw idfx to synthesized analysis).</li> <li>The selection can be explored using different interactive graph visualizations.</li> <li>Researchers can upload their own corpus (or single copy task file) and compare it to the (selected) corpus.</li> <li>An extra webpage is designed for laypersons wanting to take a copy task to test their typing skills. They get dashboard feedback in a user-friendly and attractive way and can compare their performance with (age-related) participants in the corpus. (Specially designed to further expand the corpus).</li> </ul> <p><strong>Facts and Figures</strong><br> Some facts and figures about the corpus' composition:</p> <p>Languages:</p> <ul> <li>Dutch 3130 files</li> <li>English 1163 files</li> <li>German 281 files</li> <li>French 201 files</li> <li>Other 378 file</li> </ul> <p><strong>Gender</strong></p> <ul> <li>Female: 3495 files</li> <li>Male: 1276 files</li> <li>X or missing 382 files</li> </ul> <p><strong>Age</strong></p> <ul> <li>15- 439 files</li> <li>16-20 1591 files</li> <li>21-25 2427 files</li> <li>26-35 478 files</li> <li>36-45 126 files</li> <li>46+ 230 files</li> </ul> <p>A subset of the total corpus has been uploaded here. The subset contains a dataset of about 500 tests (English | 21-25-year-olds).</p> <p> </p>
A workflow for exploring ligand dissociation from a macromolecule: Efficient random acceleration molecular dynamics simulation and interaction fingerprint analysis of ligand trajectories
<p>Containes input data for MD simulations of 3 HSP90- small compound complexes from the paper</p> <p>A workflow for exploring ligand dissociation from a macromolecule: Efficient random acceleration molecular dynamics simulation and interaction fingerprint analysis of ligand trajectories" from Daria B. Kokh, Bernd Doser , Stefan Richter , Fabian Ormersbach , Xingyi Cheng, Rebecca C. Wade, publishe in J. Chem. Phys. <strong>153</strong>, 125102 (2020); <a href="https://doi.org/10.1063/5.0019088">https://doi.org/10.1063/5.0019088</a></p> <ul> <li>ref.pdb - structure of the complex in PDB format</li> <li>ref.prmtop - topology file in AMBER</li> <li>ref-equal-NTP.pdb - structure after NTP equilibration </li> <li>ref-equal-NTP.rst7 - coordinates after NTP equilibration</li> <li>ref-equal-NTP.crd - coordinates after NTP equilibration </li> <li>gromacs.gro - coordinates in Gromacs format (after NTP equalibration)</li> <li>gromacs.top - Gromacs topology </li> </ul> <p> </p>
Data Set for the Journal Article "Autonomous Reaction Network Exploration in Homogeneous and Heterogeneous Catalysis"
<p>This dataset includes the XYZ structures of the centroids of all compounds found. Charge and multiplicity are given in the comment line of each XYZ file.</p>
How do native and non-native speakers recognize emotions in the instructor's voice in educational videos? Exploring the first step of the cognitive-affective model of e-learning for international learners [dataset]
<p>Dataset for the journal article <em>How do native and non-native speakers recognize emotions in the instructor’s voice in educational videos? Exploring the first step of the cognitive-affective model of e-learning for international learners.</em></p>
TRPC3 interaction with cholesterol as explored through MD (raw data)
<p>Transient receptor potential canonical 3 (TRPC3) channel belongs to the superfamily of transient receptor potential (TRP) channels which mediate Ca<sup>2+</sup> influx into the cell. These channels constitute essential elements of cellular signalling. TRPC3 is primarily gated by lipids, and its surface expression has been shown to be dependent on cholesterol, yet a comprehensive exploration of its interaction with this lipid has thus far not emerged. Here, through 80 µs of coarse-grained molecular dynamics simulations, we show that cholesterol interacts with multiple elements of the transmembrane machinery of TRPC3. Through our approach, we identify an annular binding site for cholesterol on the pre-S1 helix, and a non-annular site at the interface between the voltage-sensor like domain and pore domains. Here cholesterol interacts with exposed polar residues, and possibly acts to stabilise the domain interface.</p> <p> </p> <p><br> p { margin-bottom: 0.08in; color: #000000; line-height: 0.24in; text-align: justify; orphans: 2; widows: 2; background: transparent }p.western { font-family: "Palatino Linotype", serif; font-size: 12pt }p.cjk { font-family: "Palatino Linotype", serif; font-size: 12pt; so-language: de-DE }p.ctl { font-family: "Palatino Linotype", serif }a:visited { color: #954f72; text-decoration: underline }a:link { color: #0000ff; text-decoration: underline }</p> <p> </p>
Benchmark for deterministic traffic simulator - parameter space exploration (Prague, Jun 6 2021)
<p>The benchmark is meant for deterministic traffic simulator for optimising traffic flow within a city. The simulator is one part of a traffic modeling framework for intelligent transportation in smart cities. In contrast to standard navigation systems where the navigation is optimised for drivers, we aim to optimise a distribution of the global traffic flow. We utilise HPC resources for the simulator’s parameters exploration for which EVEREST SDK is used.</p> <p>The traffic simulator is available at: <a href="https://github.com/It4innovations/ruth">github.com/It4innovations/ruth</a><strong>.</strong></p> <p><br> The benchmark contains input data, routing map, and skript to run it with HyperQueue. Simulator v1.0 was used.</p>
Dataset linking to the paper "Exploring characteristics of national forest inventories for integration with global space-based forest biomass data"
<p>The dataset links to the study titled “Exploring characteristics of national forest inventories for integration with global space-based forest biomass data”. This study is published in the journal “Science of the Total Environment” and the publication can be found at <a href="https://doi.org/10.1016/j.scitotenv.2022.157788">https://doi.org/10.1016/j.scitotenv.2022.157788</a>. The dataset contains four csv files that were used to produce the results and other figures in the paper. The description of the individual data files contained in the dataset is given below.</p> <p><strong>NFI availability and characteristics data: </strong>The data file “NFI_availability_characteristics.csv” contains data on the total number of NFIs, the NFI extent, and the year of the most recent NFI in countries with NFI as reported in FRA 2020 country reports. The respective data variables in the data file are termed as Number_of_NFI, Latest_NFI_extent_FRA2020, and Latest_NFI_year_FRA2020 (NFI years generally refer to the years of data collection). In addition, the data file contains data on the region and tropical domain per country. The tropical and subtropical countries were considered tropical in the analysis and interpretation of the results. These data were used to produce Figure 2 of the study. ArcMap 10.7.1 was used for this purpose. </p> <p><strong>National biomass intercomparison data: </strong>The data file “national_biomass_intercomparison.csv” contains national forest AGB data for the year 2018 from FRA 2020 and CCI Biomass product that were used in the national biomass intercomparison analysis. The total (tons) and average space-based AGB (tons/ha) are extracted directly from the CCI Biomass Map 2018 for each country included in the study. The processing is done in Python and R environments. The spatial resolution of the map is 100 m. The average FRA AGB data in tons per ha was compiled from FRA 2020 country reports. The total FRA AGB data (tons) was estimated by multiplying each country's average FRA AGB data with FRA forest area data (in ha).</p> <p>The data unit for total AGB was converted from tons to gigaton (Gt) in intercomparison analysis. The total CCI Map AGB estimates used in the analysis are termed as CCI_MAP_AGB_Gt in the data file and the average as CCI_Map_AGB_tons.ha. Similarly, the total FRA AGB data are termed as FRA_AGB_Gt and the average as FRA_AGB_ton.ha. The NFI availability and temporality were also used in intercomparison analysis and this data is termed as Latest_NFI_year_FRA2020 in the data file. The data were used to produce Figure 3 of the study in the R environment.</p> <p><strong>NFI plot design characteristics: </strong>The data file named “NFI_plot_design_characteristics.csv” contains data on variables that were used in the analysis of NFI plot designs in 46 tropical countries. This data file mainly contains the data that was used to produce Figure 4 and Figure 6 in the R environment. The value “uniform” in the sampling_stratification variable means no stratification was used in the sampling design. The variable name “psu” stands for primary sampling unit (both cluster and single plots), “psu_distance_km” for the distance between primary sampling units in km, “cluster_plotdis_m” for the distance between plots in meter in the cluster, “plotsize_ha” for plot (single and cluster plots ) size in ha, “plotshape” for plot shapes (single and cluster plots), “ILUA” for Integrated Land Use Assessment. The data were compiled from the latest NFI design manuals and NFI reports.</p> <p><strong>NFI years: </strong>The data file “NFI_years_tropical_countries_data.csv” contains data on NFI years of the latest NFI in 46 tropical countries that were used to produce Figure 1 using ArcMap 10.7.1. The years generally refer to the last years of data collection. Data were compiled from the latest country NFI design manual or NFI report. This included both ongoing and completed NFI.</p>
Outputs of the Jupyter Notebook - Exploring Land Cover Data (Impact Observatory)
<p>The dataset contains the outputs of the notebook "Exploring Land Cover Data (Impact Observatory)" published in The Environmental Data Science Book.</p>
Exploration of historical mining site - Akersberg silver mines (St. Hanshaugen, Oslo, Norway, 04/11/2024)
<p>Exploration of the historical mining site of Akesrberg (St. Hanshaugen, Oslo, Norway, 04/11/2024)</p> <p>- Main ore minerals: pyrite, sphalerite, galena, chalcopyrite, argentite, stembergite</p> <p>- Provisional References:</p> <ul> <li>https://www.researchgate.net/publication/330971315_Akersberg_gruver_Akersberg_Silver_Mines_pages_168-174_in_Arnesen_R_Gjemte_og_glemte_steder_-_urban_utforsking_i_Oslo_og_omradet_rundt_In_Norwegian</li> <li>https://foreninger.uio.no/ngf/ngt/pdfs/NGT_71_2_121-128.pdf</li> <li>https://www.mindat.org/loc-37117.html</li> <li>https://no.wikipedia.org/wiki/Akersberg_gruver</li> </ul>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.