Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

58

datasets available to search

ShareScore release 0.7.1

Reset

Dataset results

58 results for “mapping database”

Learn how ShareScore rates datasets ↗
zenodo48/100

Oil and Gas Infrastructure Mapping (OGIM) database

<p>The Oil and Gas Infrastructure Mapping (OGIM) database is a global, spatially explicit, and granular dataset of oil and gas infrastructure. It is developed by Environmental Defense Fund (EDF)&nbsp;(<a href="https://www.edf.org/">www.edf.org</a>) and MethaneSAT, LLC (<a href="https://www.methanesat.org/">www.methanesat.org</a>), a wholly owned subsidiary of EDF. The OGIM database helps fill a crucial geospatial data need, by supporting the quantification and source characterization of oil and gas methane emissions. The database is developed via acquisition, analysis, curation, integration, and quality-assurance (performed at EDF) of publicly available geospatial data sources. These oil and gas facility datasets are reported by governments, industry, academics, and other non-government entities.</p> <p>OGIM is a collection of data tables within a GeoPackage. Each data table within the GeoPackage includes locations and facility attributes of oil and gas infrastructure types that are important sources of methane emissions, including: oil and gas production wells, offshore production platforms, natural gas compressor stations, oil and natural gas processing facilities, liquefied natural gas facilities, crude oil refineries, and pipelines. OGIM v2.7 includes approximately 6.7 million features, including 4.5 million point locations of oil and gas wells and over 1.2 million kilometers of oil and gas pipelines.</p> <p>Please see the PDF document in the &ldquo;Files&rdquo; section of this page for more information about this version, including attribute column definitions, key changes since the previous version, and more. Full details on database development and related analytics can be found in the following Earth System Science Data (ESSD) journal paper. Please cite this paper when using any version of the database:</p> <p><span>Omara, M., Gautam, R., O'Brien, M., Himmelberger, A., Franco, A., Meisenhelder, K., Hauser, G., Lyon, D., Chulakadabba, A., Miller, C., Franklin, J., Wofsy, S., and Hamburg, S.: Developing a spatially explicit global oil and gas infrastructure database for characterizing methane emission sources at high resolution, Earth Syst. Sci. Data Discuss.,&nbsp;</span><a href="https://doi.org/10.5194/essd-15-3761-2023"><span>https://doi.org/10.5194/essd-15-3761-2023</span></a><span>, 2023.</span></p> <p>Important note: While the results section of this manuscript is specific to v1 of the OGIM, the methods described therein are the same methods used to develop and update v2.7. Additionally, while we describe our data sources in detail in the manuscript above, and include maps of all acquired datasets, this open-access version of the OGIM database does not include the locations of about 300 natural gas compressor stations in Russia. Future updates may include these locations when appropriate permissions to make them publicly accessible are obtained.&nbsp;</p> <p>OGIM v2.7 is based on public-domain datasets reported in February 2025 or prior. Each record in OGIM indicates a date (SRC_DATE) when the original source of the record was published or last updated. Some records may contain out-of-date information, for example, if a facility&rsquo;s status has changed since we last visited a data source. We anticipate updating the OGIM database on a regular cadence and are continually including new public domain datasets as they become available.</p> <p>---</p> <p>Point of Contact at Environmental Defense Fund and MethaneSAT, LLC: Madeleine O&rsquo;Brien (maobrien@methanesat.org) and Mark Omara (momara@edf.org).</p>

opencc-by-4.0Jun 2024View details →
zenodo48/100

LIPID MAPS® Structure Database (LMSD) formatted for MetFrag

<p>This repository contains the LIPID MAPS&reg; Structure Database (<a href="https://www.lipidmaps.org/databases/lmsd/overview">LMSD</a>) formatted for use in <a href="https://msbi.ipb-halle.de/MetFrag/">MetFrag</a> (and other workflows).</p> <p><em>LIPID MAPS&reg; Lipidomics Gateway is a free, comprehensive website for researchers interested in lipid biology. Use <a href="https://www.lipidmaps.org"> https://www.lipidmaps.org</a> to stay abreast of developments each month from across the field, and explore the rich information collections, tools and resources from the LIPID Metabolites And Pathways Strategy (LIPID MAPS&reg;) Consortium. </em><br> &nbsp;</p> <p>The workflow used to create this file (by B. Talavera And&uacute;jar) can be found here: <a href="https://gitlab.lcsb.uni.lu/eci/simple-utilities/sdf2csv">https://gitlab.lcsb.uni.lu/eci/simple-utilities/sdf2csv</a></p> <p><strong>Reference:</strong> LMSD: LIPID MAPS&reg; structure database, Sud M., Fahy E., Cotter D., Brown A., Dennis E., Glass C., Murphy R., Raetz C., Russell D., and Subramaniam S., Nucleic Acids Research, 2006, DOI: <a href="https://doi.org/10.1093/nar/gkl838"> 10.1093/nar/gkl838 </a></p>

opencc-by-4.0Jul 2023View details →
zenodo44/100

DigiMedFor Forest Management Map Database

<p>Geodatabase of Forest Management Map created in DigiMedFor project.</p>

opencc-by-4.0May 2024View details →
zenodo44/100

SeMRA Gene Mappings Database

<p>Analyze the landscape of gene nomenclature resources, species-agnostic. See instructions for reproduction and usage in the attached README.md.</p>

opencc-zeroApr 2024View details →
zenodo44/100

SeMRA Protein Complex Mappings Database

<p>Analyze the landscape of protein complex nomenclature resources, species-agnostic. See instructions for reproduction and usage in the attached README.md.</p>

opencc-zeroApr 2024View details →
zenodo44/100

SeMRA Anatomy Mappings Database

<p>Supports the analysis of the landscape of anatomy nomenclature resources. See instructions for reproduction and usage in the attached README.md.</p>

opencc-zeroApr 2024View details →
zenodo44/100

SeMRA Cell and Cell Line Mappings Database

<p>Originally a reproduction of the EFO/Cellosaurus/DepMap/CCLE scenario posed in the Biomappings paper, this configuration imports several different cell and cell line resources and identifies mappings between them. See instructions for reproduction and usage in the attached README.md.</p>

opencc-zeroApr 2024View details →
zenodo44/100

SeMRA Disease Mappings Database

<p>Supports the analysis of the landscape of disease nomenclature resources. See instructions for reproduction and usage in the attached README.md.</p>

opencc-zeroApr 2024View details →
zenodo44/100

Gene/Protein BridgeDb ID Mapping Database (Ensembl Fungi 49)

<p>Ensembl Fungi 49 derived ID mapping database for use with BridgeDb.<br> The&nbsp;scripts used to create these databases based on Ensembl BioMart&nbsp;can be found at <a href="https://github.com/bridgedb/create-bridgedb-genedb">https://github.com/bridgedb/create-bridgedb-genedb</a>.</p> <p>This work was funded by the&nbsp;<a href="https://fairplus-project.eu/">FAIRplus project</a>&nbsp;(grant&nbsp;agreement no 802750) and&nbsp;<a href="https://www.nwo.nl/en/researchprogrammes/open-science/open-science-fund/open-science-fund-2021-awarded-grants">NWO Open Science Fund</a>&nbsp;(grant no&nbsp;<a href="https://www.nwo.nl/en/projects/203001121">203.001.121</a>).</p>

opencc-by-4.0Apr 2022View details →
zenodo44/100

Public database of geological-paleontological mapping in the surroundings of Vălioara

<p>The database contains the coordinates of the geological-paleontological mapping sites and measurements in the area of V<span>ă</span>lioara (Romania) from 2019 onwards. The Excel format file data tables contain in separate worksheets the localities, the measurements and the explanation of the mapping units. The coordinates are given in UTM34 coordinate system and also with latitude-longitude data (WGS84 datum).</p>

opencc-by-4.0Aug 2024View details →
zenodo44/100

Database for GWAS SVatalog: a visualization tool to aid fine-mapping of GWAS loci with structural variations.

<p>GWAS SVatalog is a novel visualization tool and database for structural variants (SV) found in a predominantly European population of 101 individuals with Cystic Fibrosis (CF). Aside from the CF-causing variants on chromosome 7 and the LD block in which they lie, the remainder of the genome is comparable to a the 1000 Genomes healthy European population. This data is a collection of SV calls and their linkage disequilibrium (LD) statistics with GWAS-significant SNPs reported in the GWAS Catalog.</p> <p>&nbsp;</p> <p>The goal of this project is to provide a resource to aid fine mapping of GWAS loci using SVs. GWAS loci are generally identified by SNPs which&nbsp;account for an incomplete proportion of genetic variation and phenotypic heritability.&nbsp;Their relevance to the phenotype might be limited, tagging other polymorphisms, such as SVs, that could be the cause of the association signal. To leverage this data to its full potential, visit the&nbsp;<a href="https://svatalog.research.sickkids.ca/" target="_blank" rel="noopener">GWAS SVatalog</a> web tool. Here, interactive visualizations can illustrate SVs identified in high LD with GWAS-significant SNPs, suggesting putative causal variation that could guide additional functional investigation.</p> <p>&nbsp;</p> <p>For more information on how to use GWAS SVatalog, visit the<a href="https://gwas-svatalog-docs.readthedocs.io/en/latest/index.html" target="_blank" rel="noopener noreferrer">&nbsp;documentation</a>.</p> <p>&nbsp;</p> <p>This project was accomplished in collaboration with the <a href="https://lab.research.sickkids.ca/strug/" target="_blank" rel="noopener">Strug Lab</a> at <a href="https://www.sickkids.ca/en/" target="_blank" rel="noopener">The Hospital for Sick Children (SickKids)</a>,&nbsp;<a href="https://www.tcag.ca/" target="_blank" rel="noopener">The Center for Applied Genomics (TCAG)</a>, and <a href="https://www.utoronto.ca/" target="_blank" rel="noopener">University of Toronto</a>.</p>

opencc-by-4.0Jun 2024View details →
zenodo40/100

SeMRA Raw Semantic Mappings Database

<p>An automatically assembled dataset of raw semantic mappings produced by <code>python -m semra.database</code>. This incorporates mappings from the following places:</p> <ol> <li>Ontologies indexed in the Bioregistry (primary)</li> <li>Databases integrated in PyOBO (primary)</li> <li>Biomappings (secondary)</li> <li>Wikidata (primary/secondary)</li> <li>Custom resources integrated in SeMRA (primary)</li> </ol> <p>This is a database of raw mapping without further processing. For processed mapping datasets, we suggest smaller domain-specific processing rules (see&nbsp;<a href="https://github.com/biopragmatics/semra/tree/main/notebooks/landscape">https://github.com/biopragmatics/semra/tree/main/notebooks/landscape</a> for examples). It can be accessed directly via:</p> <ul> <li><code>mappings.sssom.tsv.gz</code> - loadable through any tools supporting SSSOM</li> <li><code>mappings.jsonl.gz</code> - loadable through SeMRA using <a href="https://semra.readthedocs.io/en/latest/api/semra.io.from_jsonl.html" target="_blank" rel="noopener"><code>semra.from_jsonl</code></a></li> </ul> <h2>How to Run the Web App</h2> <ol> <li>Download all artifacts from this Record</li> <li>Make sure that you have Docker running locally</li> <li>Run <code>sh run_on_docker.sh</code> from the command line</li> <li>Navigate to http://localhost:8773 to see the SeMRA dashboard or to http://localhost:7474 for direct access to the Neo4j graph database</li> </ol> <h2>Licensing</h2> <p>Mappings are licensed according to their primary resources. These are explicitly annotated in the SSSOM file on each row (when available) and on the mapping set level in the Neo4j graph database artifacts.</p>

opencc-zeroApr 2024View details →
zenodo40/100

Geographic range maps for Mammal Diversity Database v1.3 taxonomy

<p>Update of mammal maps based on the taxonomy of the Mammal diversity database. These maps are different from the original, have been downscaled and are distributed under the R package mdd (github.com/alrobles/mdd).</p>

opencc-by-4.0Apr 2017View details →
zenodo40/100

GRASS GIS database for CASAS-PBDM (www.casasglobal.org) geospatial mapping and analysis

<p>GRASS GIS database for geospatial mapping and analysis of physiologically based demographic modeling (PBDM) implemented by the Center for the Analysis of Sustainable Agricultural Systems (CASAS,&nbsp;<a href="https://www.casasglobal.org/" target="_blank" rel="noopener">www.casasglobal.org</a>).</p> <p>The&nbsp;<code>casas_gis_grass8data.zip</code>&nbsp;archive includes data updated for use with GRASS GIS version 8.</p>

opencc-by-sa-4.0Nov 2024View details →
zenodo40/100

Gene/Protein BridgeDb ID Mapping Database (Ensembl 103)

<p>Ensembl 103 derived ID mapping database for use with BridgeDb.</p> <p><br> This work was funded by the&nbsp;<a href="https://fairplus-project.eu/">FAIRplus project</a>&nbsp;(grant&nbsp;agreement no 802750) and&nbsp;<a href="https://www.nwo.nl/en/researchprogrammes/open-science/open-science-fund/open-science-fund-2021-awarded-grants">NWO Open Science Fund</a>&nbsp;(grant no&nbsp;<a href="https://www.nwo.nl/en/projects/203001121">203.001.121</a>).</p>

openother-openFeb 2022View details →
zenodo40/100

Gene/Protein BridgeDb ID Mapping Database (Ensembl 104)

<p>Ensembl 104 derived ID mapping database for use with BridgeDb.</p> <p>This work was funded by the&nbsp;<a href="https://fairplus-project.eu/">FAIRplus project</a>&nbsp;(grant&nbsp;agreement no 802750) and&nbsp;<a href="https://www.nwo.nl/en/researchprogrammes/open-science/open-science-fund/open-science-fund-2021-awarded-grants">NWO Open Science Fund</a>&nbsp;(grant no&nbsp;<a href="https://www.nwo.nl/en/projects/203001121">203.001.121</a>).</p>

openother-openMar 2022View details →
zenodo40/100

Gene/Protein BridgeDb ID Mapping Database (Ensembl 105)

<p>Ensembl 105 derived ID mapping databases for use with BridgeDb.</p> <p>The&nbsp;scripts used to create these databases based on Ensembl BioMart&nbsp;can be found at <a href="https://github.com/bridgedb/create-bridgedb-genedb">https://github.com/bridgedb/create-bridgedb-genedb</a>.</p> <p>This work was funded by the&nbsp;<a href="https://fairplus-project.eu/">FAIRplus project</a>&nbsp;(grant&nbsp;agreement no 802750) and&nbsp;<a href="https://www.nwo.nl/en/researchprogrammes/open-science/open-science-fund/open-science-fund-2021-awarded-grants">NWO Open Science Fund</a>&nbsp;(grant no&nbsp;<a href="https://www.nwo.nl/en/projects/203001121">203.001.121</a>).</p>

openother-openApr 2022View details →
zenodo40/100

Gene/Protein BridgeDb ID Mapping Database (Ensembl Metazoa 49)

<p>Ensembl Metazoa 49 derived ID mapping databases for use with BridgeDb.<br> The&nbsp;scripts used to create these databases based on Ensembl BioMart&nbsp;can be found at <a href="https://github.com/bridgedb/create-bridgedb-genedb">https://github.com/bridgedb/create-bridgedb-genedb</a>.</p> <p>This work was funded by the&nbsp;<a href="https://fairplus-project.eu/">FAIRplus project</a>&nbsp;(grant&nbsp;agreement no 802750) and&nbsp;<a href="https://www.nwo.nl/en/researchprogrammes/open-science/open-science-fund/open-science-fund-2021-awarded-grants">NWO Open Science Fund</a>&nbsp;(grant no&nbsp;<a href="https://www.nwo.nl/en/projects/203001121">203.001.121</a>).<br> &nbsp;</p>

openother-openApr 2022View details →
zenodo40/100

Gene/Protein BridgeDb ID Mapping Database (Ensembl Plants 49)

<p>Ensembl Plants 49 derived ID mapping database for use with BridgeDb.<br> The&nbsp;scripts used to create these databases based on Ensembl BioMart&nbsp;can be found at <a href="https://github.com/bridgedb/create-bridgedb-genedb">https://github.com/bridgedb/create-bridgedb-genedb</a>.</p> <p>This work was funded by the&nbsp;<a href="https://fairplus-project.eu/">FAIRplus project</a>&nbsp;(grant&nbsp;agreement no 802750) and&nbsp;<a href="https://www.nwo.nl/en/researchprogrammes/open-science/open-science-fund/open-science-fund-2021-awarded-grants">NWO Open Science Fund</a>&nbsp;(grant no&nbsp;<a href="https://www.nwo.nl/en/projects/203001121">203.001.121</a>).</p>

openother-openApr 2022View details →
zenodo40/100

Derby database for mapping secondary to primary HMDB identifiers

<p>The data (hmdb_metabolites, released on 17/11/2021) used to create this ID mapping database was downloaded from HMDB (<em>Human Metabolome Database,&nbsp;</em>website URL:&nbsp;https://hmdb.ca/).&nbsp;</p> <p>This database was used for the <a href="https://github.com/tabbassidaloii/BridgeDbDemoBioSB2022">BridgeDb demo at BioSB 2022</a> conference.</p> <p>The&nbsp;scripts used to create this&nbsp;database&nbsp;based on HGNC: https://github.com/tabbassidaloii/create-bridgedb-secondary2primary</p> <p>This work was funded by the&nbsp;<a href="https://fairplus-project.eu/">FAIRplus project</a>&nbsp;(grant&nbsp;agreement no 802750) and&nbsp;<a href="https://www.nwo.nl/en/researchprogrammes/open-science/open-science-fund/open-science-fund-2021-awarded-grants">NWO Open Science Fund</a>&nbsp;(grant no&nbsp;<a href="https://www.nwo.nl/en/projects/203001121">203.001.121</a>).</p>

opencc-by-4.0Jun 2022View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record