Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
468
datasets available to search
ShareScore release 0.9.0
Dataset results
468 results for “dump”
Wikidata Dump judges-de
<p>RDF dump of wikidata produced with <a href="//wdumps.toolforge.org/">wdumper</a>.</p><p><br><a href="//wdumps.toolforge.org/dump/2892">View on wdumper</a></p><p><b>entity count<b>: 3845, <b>statement count</b>: 88346, <b>triple count</b>: 879361</b></b></p>
Wikidata Dump Wikidata_human
<p>RDF dump of wikidata produced with <a href="//wdumps.toolforge.org/">wdumper</a>.</p><p><br><a href="//wdumps.toolforge.org/dump/2929">View on wdumper</a></p><p><b>entity count<b>: 0, <b>statement count</b>: 0, <b>triple count</b>: 38</b></b></p>
Wikidata Dump test
<p>RDF dump of wikidata produced with <a href="//wdumps.toolforge.org/">wdumper</a>.</p><p><br><a href="//wdumps.toolforge.org/dump/2920">View on wdumper</a></p><p><b>entity count<b>: 0, <b>statement count</b>: 0, <b>triple count</b>: 0</b></b></p>
QALD-10 Wikidata Dump
<p>This data dump of Wikidata is published to allow fair and replicable evaluation of KGQA systems with the <a href="https://github.com/KGQA/QALD_10">QALD-10 benchmark</a>. QALD-10 is newly released and was used in the <a href="https://www.nliwod.org/challenge">QALD-10 Challenge</a>. Anyone interested in evaluating their KGQA systems with QALD-10 can download this dump and set up a local Wikidata endpoint in their server.</p>
Wikidata Dump religions-wikidata
<p>RDF dump of wikidata produced with <a href="//wdumps.toolforge.org/">wdumper</a>.</p><p>religions wikidata<br><a href="//wdumps.toolforge.org/dump/2978">View on wdumper</a></p><p><b>entity count<b>: 0, <b>statement count</b>: 0, <b>triple count</b>: 0</b></b></p>
Wikidata Dump religions-wikidata
<p>RDF dump of wikidata produced with <a href="//wdumps.toolforge.org/">wdumper</a>.</p><p>religions wikidata<br><a href="//wdumps.toolforge.org/dump/2978">View on wdumper</a></p><p><b>entity count<b>: 0, <b>statement count</b>: 0, <b>triple count</b>: 0</b></b></p>
Wikidata Dump religions-wikidata
<p>RDF dump of wikidata produced with <a href="//wdumps.toolforge.org/">wdumper</a>.</p><p>religions wikidata<br><a href="//wdumps.toolforge.org/dump/2978">View on wdumper</a></p><p><b>entity count<b>: 0, <b>statement count</b>: 0, <b>triple count</b>: 0</b></b></p>
Wikidata Dump religions-wikidata
<p>RDF dump of wikidata produced with <a href="//wdumps.toolforge.org/">wdumper</a>.</p><p>religions wikidata<br><a href="//wdumps.toolforge.org/dump/2978">View on wdumper</a></p><p><b>entity count<b>: 0, <b>statement count</b>: 0, <b>triple count</b>: 0</b></b></p>
Raw strain and temperature data for a dump truck, IFV, and semi-truck
<p>Horizontal strain measurements for a dump truck (D), an IFV (K), and a semi-truck (L). The measurements were taken from DTU smart road, which consists of four sections, each with six horizontal strain gauges placed at the bottom of the AC (150 mm depth). There is a reinforcement grid in section 1, 2, and 3. The data was taken as a part of the Master thesis: Investigation of asphalt pavements loaded by heavy off-road vehicles. The thesis contains further details.</p>
Open Context Database SQL Dump: Legacy Schema Tables and New Schema Tables
<p>Open Context (<a href="https://opencontext.org">https://opencontext.org</a>) publishes free and open access research data for archaeology and related disciplines. An open source (but bespoke) Django (Python) application supports these data publishing services. The software repository is here: <a href="https://github.com/ekansa/open-context-py">https://github.com/ekansa/open-context-py</a></p> <p>The Open Context team runs ETL (extract, transform, load) workflows to import data contributed by researchers from various source relational databases and spreadsheets. Open Context uses PostgreSQL (<a href="https://www.postgresql.org">https://www.postgresql.org</a>) relational database to manage these imported data in a graph style schema. The Open Context Python application interacts with the PostgreSQL database via the Django Object-Relational-Model (ORM).</p> <p>In 2023, the Open Context team finished migration of from a legacy database schema to a revised and refactored database schema with stricter referential integrity and better consistency across tables. During this process, the Open Context team de-duplicated records, cleaned some metadata, and redacted attribute data left over from records that had been incompletely deleted in the legacy schema.</p> <p>This database dump includes all Open Context data organized with the legacy schema (table names that start with the 'oc_' or 'link_' prefixes) along with all Open Context data after cleanup and migration to the new database schema (table names that start with 'oc_all_'). The binary media files referenced by these structured data records are stored elsewhere. Binary media files for some projects, still in preparation, are not yet archived with long term digital repositories.</p> <p>These data comprehensively reflect the structured data currently published and publicly available on Open Context. Other data (such as user and group information) used to run the Website are not included. </p> <p> </p> <p><strong>IMPORTANT</strong></p> <p>This database dump contains data from roughly 180 different projects. Each project dataset has its own metadata and citation expectations. If you use these data, you must cite each data contributor appropriately, not just this Zenodo archived database dump.</p> <p> </p> <p> </p>
Scholix dump of the OpenAIRE inferred citations
<p>This dataset contains the set of citations extracted for the large by the OpenAIRE Information Inference Service (IIS). It consists of tar archives, each containing gzip-compressed files, including new-line delimited JSON records in Scholix format.</p> <p>The dataset counts 36.430.057 citations, from citing scientific products.</p> <p>The type of PIDs among the citing and the cited products include:</p> <ul> <li>DOI</li> <li>PMC</li> <li>PMID</li> <li>ArXiv</li> <li>Handle</li> </ul>
Generated Wikidata Subset for geneWiki based on dump: GeneWiki_wikidata-20220630
<p>Source file: GeneWiki_wikidata-20220630-all.ttl</p>
Generated Wikidata Subset for geneWiki based on dump: wikidata-20150601-all
<p>Source file: wikidata-20150601-all.ttl</p>
Generated Wikidata Subset for geneWiki based on dump: GeneWiki_wikidata-20210531-all
<p>Source file: GeneWiki_wikidata-20210531-all.ttl.gz</p>
Generated Wikidata Subset for geneWiki based on dump: GeneWiki_wikidata-20180115-all
<p>Source file: GeneWiki_wikidata-20180115-all.ttl.gz</p>
Generated Wikidata Subset for geneWiki based on dump: GeneWiki_wikidata-20190121-all
<p>Source file: GeneWiki_wikidata-20190121-all.ttl.gz</p>
Generated Wikidata Subset for geneWiki based on dump: GeneWiki_wikidata-20170821-all
<p>Source file: GeneWiki_wikidata-20170821-all.ttl.gz</p>
Generated Wikidata Subset for geneWiki based on dump: GeneWiki_wikidata-20201102-all
<p>Source file: GeneWiki_wikidata-20201102-all.ttl.gz</p>
Generated Wikidata Subset for geneWiki based on dump: GeneWiki_wikidata-20160613-all
<p>Source file: GeneWiki_wikidata-20160613-all.ttl.gz</p>
Wikidata Dump Swedish places
<p>RDF dump of wikidata produced with <a href="//wdumps.toolforge.org/">wdumper</a>.</p><p>Swedish places with coordinates<br><a href="//wdumps.toolforge.org/dump/3241">View on wdumper</a></p><p><b>entity count<b>: 9813372, <b>statement count</b>: 91203322, <b>triple count</b>: 100921154</b></b></p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.