Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

1,481

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

1,481 results for “data processing”

Learn how ShareScore rates datasets ↗
zenodo32/100

Data for "Stable climate simulations using a realistic GCM with neural network parameterizations for atmospheric moist physics and radiation processes"

<p>This is&nbsp;sampling data&nbsp;of &quot;Stable climate simulations using a realistic GCM with neural network parameterizations for atmospheric moist physics and radiation processes&quot;.</p> <p>&#39;qv_nn_in&#39; for the large scale specific humidity, [kg/kg]<br> &#39;T_nn_in&#39; for the large scale temperature, [K]<br> &#39;dqvls_nn_in&#39; for the large scale moisture advection, [kg/kg/s]<br> &#39;dTls_nn_in&#39; for the large scale moisture advection, [K/s]<br> &#39;qtend_check&#39; for the moistening rate by CRM, [kg/kg/s]<br> &#39;stend_check&#39; for the heating rate by CRM, [K/s]<br> &#39;SOLS&#39; for direct shorwave solar radiation down to surface, [W/m2]<br> &#39;SOLSD&#39; for diffusive shortwave solar radiation down to surface, [W/m2]<br> &#39;SOLL&#39; for direct near infrared solar radiation down to surface, [W/m2]<br> &#39;SOLLD&#39; for diffusive near infrared solar radiation down to surface, [W/m2]<br> &#39;SOLIN&#39; for insolation at model top, [W/m2]<br> &#39;FSNS&#39; for net shortwave radiation at model surface, [W/m2]<br> &#39;FSNT&#39; for net shortwave radiation at model top, [W/m2]<br> &#39;FLNS&#39; for net longwave radiation at model surface, [W/m2]<br> &#39;FLNT&#39; for net longwave radiation at model top, [W/m2]<br> &#39;SPPS&#39; for surface pressure, [Pa]</p> <p>To download the full dataset of the SPCAM simulation in 1998. Please click the dropbox link: https://www.dropbox.com/s/p841v1tw00rokdy/SPCAM_VAR_1998.tar.gz?dl=0</p>

opencc-by-4.0Oct 2021View details →
zenodo32/100

Processed data for the "Deriving Semantics-Aware Fuzzers from Web API Schemas" paper

<p>Processed data for the &quot;Deriving Semantics-Aware Fuzzers from Web API Schemas&quot; paper.&nbsp; Each directory in the archive consists of:</p> <p>- metadata.json. Metadata about a test run - tested fuzzer name, run duration, etc</p> <p>- fuzzer.json&nbsp;- Structured fuzzer output</p> <p>-&nbsp;deduplicated_cases.json - Deduplicated reported failures, when fuzzers provide it</p> <p>- sentry.json&nbsp;- Cleaned Sentry events for this run</p> <p>- target.json&nbsp;- Parsed stdout for Gitlab &amp; Disease.sh targets that were tested without Sentry integration</p>

opencc-by-4.0Nov 2021View details →
zenodo32/100

Processed Data and codes for visualization

<p>Processed Data and codes for visualization</p>

opencc-by-4.0Nov 2021View details →
dryad32/100

Forest spider data to study the response of assemblages to regional and local processes

<p>Understanding species richness variation among local communities is one of the central topics in ecology, but the complex interplay of regional processes, environmental filtering and local processes hampers generalization on the importance of different processes. Here, we aim to  unravel drivers of spider community assembly in temperate forests by analyzing two independent data sets covering gradients in elevation and forest succession. We test the following four hypotheses: (H1) Spider assemblages within a region are limited by dispersal; (H2) Local environment has a disproportionate influence on species composition; and (H3) resources and (H4) biotic interactions both affect species richness patterns. Spider data were collected by pitfall traps in two independent projects within a region along elevation and canopy cover gradients.</p>

opencc-zeroDec 2021View details →
zenodo32/100

rOMT processed data: Cerebral amyloid angiopathy is associated with glymphatic transport reduction and time-delayed solute drainage along the neck arteries

<p>This dataset contains the speed map and P&eacute;clet map data processed by regularized optimal mass transport method (<a href="https://zenodo.org/record/5809635#.YczuJy2ZNBw">https://zenodo.org/record/5809635#.YczuJy2ZNBw</a>,&nbsp;<a href="https://github.com/xinan-nancy-chen/rOMT">https://github.com/xinan-nancy-chen/rOMT</a>).</p> <p>All files are in nifty format. This dataset contains in total 55 rat cases which are divided by age (3-month, 6-month and 12-month) and status of health (&quot;WT&quot; for wild-type and &quot;CAA&quot; for those who have developed Cerebral Amyloid Angiopathy).</p>

opencc-by-4.0Dec 2021View details →
zenodo32/100

Scanning transmission electron microscopy data of LiNi0.5Co0.2Mn0.3O2 single crystal Cathode materials during degradation process

<p>Scanning transmission electron microscopy&nbsp;data of LiNi0.5Co0.2Mn0.3O2 single-crystal Cathode materials during the degradation process</p>

opencc-by-4.0Dec 2021View details →
zenodo32/100

Selected processed 5G base station RF-EMF measurement data

<p>This presents the selected processed 5G base station (BS) radio frequency electromagnetic field (RF-EMF) measurement data acquired under measurement campaign&nbsp;in outdoor environment. This data links to the findings shown in Section 5&nbsp;of D1 report at http://empir.npl.co.uk/5grfex/wp-content/uploads/sites/55/2022/01/updated-EMPIR-18SIP02-5GRFEX-Deliverable-Report-D1.pdf.&nbsp;</p> <p>This work was supported by the EU project 5GRFEX&nbsp;entitled &ndash; &lsquo;Metrology for RF exposure from Massive MIMO&nbsp;5G base station: Impact on 5G network deployment&rsquo; (this&nbsp;project has received funding from the support for impact&nbsp;(SIP) programme co-financed by the Participating States and&nbsp;from the European Union&rsquo;s Horizon 2020 research and&nbsp;innovation programme), under European Association of&nbsp;National Metrology Institutes (EURAMET) Reference&nbsp;18SIP02.</p>

opencc-by-4.0Jan 2022View details →
zenodo32/100

TopiOCQA processed Wikipedia data

Processed Wikipedia data for TopiOCQA

opencc-zeroFeb 2022View details →
zenodo32/100

Transcriptomics processed data outputs for Sofen et al 2022

<p>Processed data ouputs for transcripomtics analysis of Southern Ocean Time Series eukaryotic microbial communities. Files include: protein and HMM&nbsp;database used for iron stress biomarker screening, Proteins that hit to the HMM database curated to search the metatranscriptomic dataset for iron stress biomarker proteins with&nbsp;protein ID&#39;s, taxonomic annotation (EUKulele, MMTESP) and calculated transcripts per million (TPM) values, DNA-directed RNA Polymerase (RPB1) proteins used for taxonomic charecterization&nbsp;of the surface eukaryotic community with&nbsp;protein ID&#39;s, taxonomic annotation (EUKulele, MMTESP) and calculated transcripts per million (TPM) values.</p>

opencc-by-4.0Feb 2022View details →
zenodo32/100

RAW Data - Mechanical characterization for the severely processed FSPed WE54 magnesium alloy

<p>The present data set present the raw data of the mechanical characterization of a WE54 magnesium alloy, processed by friction stir processing (FSP), using a refrigerated backing anvil. File names describe the kind and characteristics of each test as follows:&nbsp;</p> <ol> <li>All file names start by the initial temper of the WE54 magnesium alloy (T6 or TT) followed by the FSP processing conditions: first two digits are the rotation speed while the second two digits correspond to the advancing speed, both divided by a factor of 100.</li> <li>Files tagged at the end as _iUMI.opj correspond to the characterization by instrumented ultra-microindentation and provide a data matrix including prosition and hardness value (readable in Origin).</li> <li>Files tagged at the end as _CSRtt.opj&nbsp;correspond to the characterization by constant strain rate tensile test.</li> </ol>

opencc-by-4.0Feb 2022View details →
dryad32/100

Data from: Innovative ochre processing and tool-use in China 40,000 years ago

<p>These data were generated to determine the anthropogenic hematite grains resulted from ochre processing at the Xiamabei site in the Nihewan Basin, northern China. Four samples were selected for sediment analyses, including Raman spectroscopy, X-ray diffraction (XRD), high-temperature magnetic susceptibilities, and/or magnetic component analysis of coercivity distributions. Two sediment samples (X1 and X2) come from the red stained area on which the ochre fragments OP1 and OP2, stone slab LS and quartzite cobble QC were found; two samples (X3 and X6) were retrieved in the same layer but at ~2 m distance from the stained area. All of the samples consist of flood plain silts.</p> <p>The Raman spectra unambiguously identify hematite in red particles abundantly present in samples X1 and X2. The XRD measurements clearly show that hematite is abundant in samples X1 and X2. However, hematite is not detectable in samples X3 and X6. The volume percent for hematite is 0.9% in sample X1 and 1.7% in sample X2. The high-temperature magnetic susceptibility measurements suggest that hematite dominates the magnetic mineralogy of samples X1 and X2. Magnetic component analyses of coercivity distributions show distinct assemblages of magnetic minerals in samples X1 and X2, which have three components with low, middle and high coercivities. The low-coercivity component with median acquisition field of 40-62 mini Tesla (mT) is interpreted as magnetite and/or maghemite. The middle-coercivity component with median acquisition field of 100-166 mT is interpreted as partially-oxidized coarse-grained magnetite. The high-coercivity component with median acquisition field of up to 575 mT is interpreted as single domain hematite, because the single domain threshold grain size of hematite is considerably larger than 15 micrometers, and even up to 100 micrometers. This kind of hematite grains with high coercivities up to several hundreds of mT is usually of detrital origin, and is documented as evidence for the earliest ochre processing in east Asia. The new findings provide new insights into the expansion of Homo sapiens.</p>

opencc-zeroFeb 2022View details →
zenodo32/100

Processed data from Repair-seq screens of prime editing of point mutations

<p>Processed data from Repair-seq screens of prime editing of point mutations.</p>

opencc-by-4.0Oct 2021View details →
zenodo32/100

Processed data: flood generation processes and flood anomalies

<p>This dataset inludes processed data required to reproduce main results presented in the manuscript by Tarasova et al. &quot;Shifts in flood generation processes exacerbate regional flood anomalies in Europe&quot;.</p>

opencc-by-4.0Mar 2022View details →
zenodo32/100

Processed data for "SpotClean adjusts for spot swapping in spatial transcriptomics data"

<p>This repo contains processed data to reproduce results in the paper&nbsp;&quot;SpotClean adjusts for spot swapping in spatial transcriptomics data&quot;.</p>

opencc-by-4.0Apr 2022View details →
zenodo32/100

The input files and processed output data for the OSSEs using particle filter for the forecast of cloud and precipitation

<p>DART_input.tar.gz contains the input files for the DART system for the control run, Exp-1 ~ Exp-5.</p> <p>WRF&amp;WRS_namelist.tar.gz contains the input namelist files for the WPS and WRF model for the nature run, the control run, and Exp-1 ~ Exp-5.</p> <p>cloud_data.tar.gz contains the procssed output data for the WRF/DART system corresponding to the nature run, the control run, and Exp-1 ~ Exp-5. Untarring the file will generate several subdirectories named after the UTC time. For example, 202008200200 denotes the results for at 02:00 UTC, 20 August, 2020. Because the data for all times are quite large (~10G), we only uploaded the data at 02:00 UTC, 20 August, 2020, 10:30 UTC, 20 August, 2020, 19:00 UTC, 20 August, 2020, 03:30 UTC, 21 August, 2020, and 12:00 UTC, 21 August, 2020. Each subdirectory contains preassim_mean.nc and postassim_mean.nc, which mean the posterior and prior estimate of atmosphere state variables and other processed variables. Either preassim_mean.nc or postassim_mean.nc contains the cloud water path (CWP, unit:kgm-2), cloud water content (CWC, which is the sum of the mixing ratio of six cloud hydrometeors, unit: kgkg-1), RE_CLOUD(unit:um), RE_ICE(unit:um), QVAPOR(unit:kgkg-1), T(the perturbation of potential temperature, unit:K).</p> <p>rain_rate.tar.gz contains the procssed output data for the rain rate corresponding to the nature run, the control run, and Exp-4 ~ Exp-5.</p>

opencc-by-4.0May 2022View details →
zenodo32/100

LTEE-TnSeq-processed-data

<p>This repository contains processed data from transposon sequencing (TnSeq) and whole genome sequencing (WGS) of the ancestral strains and evolved clones at 50K from the Lenski Long-Term Evolution Experiment (LTEE). It comprises TnSeq mutant trajectory counts, and output files from WGS analysis using breseq and samtools.&nbsp;</p>

opencc-by-4.0May 2022View details →
zenodo32/100

Underlying data / Hydrothermal carbonization as an alternative sanitation technology: process optimization and development of low-cost reactor

<p>Experimental data of a research publication &quot;Hydrothermal carbonization as an alternative sanitation technology: process optimization and development of low-cost reactor&quot; is provided.</p> <p><strong>Background:</strong> The provision of safe sanitation services is essential for human well-being and environmental integrity, but it is often lacking in less developed communities with insufficient financial and technical resources. Hydrothermal carbonization (HTC) has been suggested as an alternative sanitation technology, producing value-added products from faecal waste. We evaluated the HTC technology for raw human waste treatment in terms of resource recovery. In addition, we constructed and tested a low-cost HTC reactor for its technical feasibility.</p> <p><strong>Methods: </strong>Raw human faeces were hydrothermally treated in a mild severity range (&le; 200 &deg;C and &le; 1 hr). The total energy recovery was analysed from the energy input, higher heating value (HHV) of hydrochar and biomethane potential of process water. The nutrient contents were recovered through struvite precipitation employing process water and acid leachate from hydrochar ash. A bench-scale low-cost reactor (BLR) was developed using widely available materials and tested for human faeces treatment.</p> <p><strong>Results: </strong>The hydrochar had HHVs (23.2 - 25.2 MJ / kg) comparable to bituminous coal. The calorific value of hydrochar accounted for more than 90 % of the total energy recovery. Around 78 % of phosphorus in feedstock was retained in hydrochar ash, while 15 % was in process water. 72 % of the initial phosphorus can be recovered as struvite when deficient Mg and NH<sub>4</sub> are supplemented. The experiments with BLR showed stable operation for faecal waste treatment with an energy efficiency comparable to a commercial reactor system.</p> <p><strong>Conclusions: </strong>&nbsp;This research presents a proof of concept for the hydrothermal treatment of faecal waste as an alternative sanitation technology, by providing a quantitative evaluation of the resource recovery of energy and nutrients. The experiments with the BLR demonstrate the technical feasibility of the low-cost reactor and support its further development on a larger scale to reach practical implementation</p> <p>&nbsp;</p>

opencc-by-4.0Nov 2021View details →
zenodo32/100

Processed Data 1 : Spatial multi-omic map of human myocardial infarction

<p>We provide here the processed snRNA-seq, snATAC-seq, and visium data for&nbsp;the manuscript: Kuppe, Ramirez Flores, Li et al. &quot;Spatial multi-omic map of human myocardial infarction&quot;, 2022</p>

opencc-by-4.0May 2022View details →
dryad32/100

Data from: Parallel processing in speech perception with local and global representations of linguistic context

<p>Speech processing is highly incremental. It is widely accepted that human listeners continuously use the linguistic context to anticipate upcoming concepts, words, and phonemes. However, previous evidence supports two seemingly contradictory models of how a predictive context is integrated with the bottom-up sensory input: Classic psycholinguistic paradigms suggest a two-stage process, in which acoustic input initially leads to local, context-independent representations, which are then quickly integrated with contextual constraints. This contrasts with the view that the brain constructs a single coherent, unified interpretation of the input, which fully integrates available information across representational hierarchies, and thus uses contextual constraints to modulate even the earliest sensory representations. To distinguish these hypotheses, we tested magnetoencephalography responses to continuous narrative speech for signatures of local and unified predictive models. Results provide evidence that listeners employ both types of models in parallel. Two local context models uniquely predict some part of early neural responses, one based on sublexical phoneme sequences, and one based on the phonemes in the current word alone; at the same time, even early responses to phonemes also reflect a unified model that incorporates sentence-level constraints to predict upcoming phonemes. Neural source localization places the anatomical origins of the different predictive models in nonidentical parts of the superior temporal lobes bilaterally, with the right hemisphere showing a relative preference for more local models. These results suggest that speech processing recruits both local and unified predictive models in parallel, reconciling previous disparate findings. Parallel models might make the perceptual system more robust, facilitate processing of unexpected inputs, and serve a function in language acquisition.</p>

opencc-zeroMay 2022View details →
zenodo32/100

Processed data and scripts for visualization of Kotsuki et al. (2022; GMDD)

<p>Processed data and scripts for visualization of Kotsuki et al. (2022; GMDD)</p>

opencc-by-4.0May 2022View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record