Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
403
datasets available to search
ShareScore release 0.9.0
Dataset results
403 results for “Occurrence Data”
Data from: Inferring new relations between medical entities using literature curated term co-occurrences
Open the record for dataset details and reuse information.
Data from: A multi-state dynamic occupancy model to estimate local colonization-extinction rates and patterns of co-occurrence between two or more interacting species
Open the record for dataset details and reuse information.
Data from: An automated approach to identifying search terms for systematic reviews using keyword co-occurrence networks
Open the record for dataset details and reuse information.
Data from: Molecular phylogeny of Squaliformes and first occurrence of bioluminescence in sharks
Open the record for dataset details and reuse information.
Data from: Occurrence, costs and heritability of delayed selfing in a free-living flatworm
Open the record for dataset details and reuse information.
Data from: Heterogeneous matrix habitat drives species occurrences in complex, fragmented landscapes
Open the record for dataset details and reuse information.
Data from: Separation of realized ecological niche axes among sympatric tilefishes provides insight into potential drivers of co‐occurrence in the NW Atlantic
Open the record for dataset details and reuse information.
Asimina triloba georeferenced occurrence data and genetic data
Open the record for dataset details and reuse information.
Data from: The role of human outdoor recreation in shaping patterns of grizzly bear-black bear co-occurrence
Open the record for dataset details and reuse information.
Data from: Taxon abundance, diversity, co-occurrence and network analysis of the ruminal microbiota in response to dietary changes in dairy cows
Open the record for dataset details and reuse information.
Data from: On the occurrence of false positives in tests of migration under an isolation with migration model
Open the record for dataset details and reuse information.
Environmental drivers of plant distributions at global and regional scales: occurrence data with associated environmental variables of plant families/genera/species
Open the record for dataset details and reuse information.
Data from: Global induced stress field from large earthquakes since 1900 and chained earthquake occurrence
Open the record for dataset details and reuse information.
CALIPSO Lidar Level 3 Cloud Occurrence Data, Standard V1-00
CAL_LID_L3_Cloud_Occurrence-Standard-V1-00 is the Cloud-Aerosol Lidar and Infrared Pathfinder Satellite Observation (CALIPSO) Lidar Level 3 Cloud Occurrence Data, Standard Version 1-00 data product. This data product was collected using the Cloud-Aerosol Lidar with Orthogonal Polarization (CALIOP) instrument. The degradation of the laser energies that started in September 2016 had a negative impact on the product, and because of this, generation and distribution ended in December 2016. Updated Lidar Level 2 data products and changes to the Lidar Level 3 Cloud Occurrence algorithm will need to be completed before a new release of this product is released.This product reports global distributions of clouds on a uniform spatial grid. All level 3 parameters are derived from the CALIPSO level 2 data, with a temporal average of one month. CALIPSO was launched on April 28, 2006, and continues to collect data necessary to study the impact of clouds and aerosols on the Earth's radiation budget and climate. It flies in the international A-Train constellation for coincident Earth observations. The CALIPSO satellite comprises three instruments: CALIOP, Imaging Infrared Radiometer (IIR), and Wide Field Camera (WFC). CALIPSO is a joint satellite mission between NASA and the French Agency CNES.
A data science approach to climate change risk assessment applied to pluvial flood occurrences for the United States and Canada (Supplementary Material)
<p>This resource is tied to manuscript "A data science approach to climate change risk assessment applied to pluvial flood occurrences for the United States and Canada".</p> <p>It contains: (1) a PDF with additional details on the implementation of GLM, GAM and RF, summarized outputs for two models, a bias analysis of the CRCM and extensive tables from Section 5.2; (2) full outputs for two models; (3) high-resolution figures, and; (4) rasters for selected figures.</p> <p>**</p> <p>Version 2 adds a log-log version of Figure 9 and changes the main PDF file to remove blue section titles.</p>
Data sets for "On the quiet-time occurrence rates, severity and origin of L-band ionospheric scintillations observed from low-to-mid latitude sites located in Puerto Rico" by Gomez Socola et al.
<p>These data sets contain the night-time scintillation index (S4) of the geomagnetically quiet days. </p>
Data from: Plant-pollinator interactions over 120 years: loss of species, co-occurrence, and function
Open the record for dataset details and reuse information.
Fig. 2 in New data on the occurrence of longhorn beetles (Coleoptera: Cerambycidae) in the Eastern Beskid Mountains (Poland)
Fig. 2. Pupa and larval feeding grounds of Nivellia sanguinosa (Gyllenhal, 1827) in a Padus avium (Mill.) trunk. Photo by W.T. Szczepański.
Data from: Using a null model to recognize significant co-occurrence prior to identifying candidate areas of endemism
[No abstract entered]
Data and R code for: "Nineteenth-century land use shape the current occurrence of some plant species, but weakly affects richness and total composition of Central European grasslands"'
<ol> <li> <p><strong><code>IndVal.all.habitats.csv</code></strong>: the results of the IndVal statistics (<a href="https://doi.org/10.1111/j.1600-0706.2010.18334.x">De Cáceres et al. 2013</a>) for 1,498 species for the historical land use categories calculated across the entire dataset;</p> </li> <li> <p><code><strong>IndVal.separate.habitats.csv</strong></code>: the results of the IndVal statistics for 1,498 species for the historical land use categories calculated for each habitat type (dry grasslands, mesic grasslands, wet grasslands) separately;</p> </li> <li> <p><code><strong>ecological.and.disturbance.values.csv</strong></code>: the original Ellenberg-type and disturbance indicator values, and the varimax-rotated components (‘RC’) used in the analysis (data obtained from <a href="https://doi.org/10.1111/jvs.13168">Tichý et al. 2023</a> and <a href="http://dx.doi.org/10.1111/geb.13603">Midolo et al. 2023</a>; accessible at the FloraVeg.eu website <a href="https://floraveg.eu/download/" target="_new" rel="noreferrer">https://floraveg.eu/download/</a>);</p> </li> <li> <p><strong>R code and data for reproducibility</strong>. The R code is for illustration purposes only and is based on a subset of 1,184 mesic grassland vegetation plots located in the Czech Republic and in the study area. This is part of the Czech National Phytosociological Database (<a href="https://www.preslia.cz/article/387">Chytrý & Rafajová 2003</a>) and the European Vegetation Archive (<a href="https://doi.org/10.1111/avsc.12191">Chytrý et al. 2016</a>). The data includes the following:</p> <ul> <li> <p> <code>data</code> folder:</p> </li> </ul> </li> </ol> <ul> <li> <ul> <li> <ul> <li>i. <code>indicator.values.csv</code>: the original indicator values for 831 species;</li> <li>ii. <code>plot.data.csv</code>: data for each of the 1,184 vegetation plots, including their historical land use, plot size, bioclimatic variables (‘bio’; <a href="http://dx.doi.org/10.1038/sdata.2017.122">Karger et al. 2017</a>), and soil pH (<a href="https://doi.org/10.1371%2Fjournal.pone.0169748">Hengl et al. 2017</a>);</li> <li>iii. <code>species.matrix.csv</code>: community matrix reporting the relative abundance of species (columns) and plot sites (rows).</li> </ul> </li> <li>R scripts for species richness, species composition, and species indicator analyses. R script are also rendered in .html with R Markdown.</li> </ul> </li> </ul>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.