Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
100
datasets available to search
ShareScore release 0.9.0
Dataset results
100 results for “rdf”
ONSTAGE as RDF triples
<p>ONSTAGE is the online data repository for Amsterdam’s first and most prominent public theatre venue: the <em>Schouwburg</em>. Currenty the database contains all programs from 1638 until 1940. All data is available on the Web as Linked Open Data in the RDFa format. The file stored here, contains all data in the N-triples RDF serialization format.</p>
RDF models and SPARQL queries for decoupled analytics demonstration
<p>This directory contains the following:</p> <p>- Two sets of RDF triples using different ontologies, modeling the same "MZVAV-2" air handling unit from the data inventory by Granderson et al. [1].</p> <p>- SPARQL queries for retrieving inputs to APAR [2] rules.</p> <p>- SPARQL queries for discovering "data links" from the models, connecting data points to time series providers.</p> <p>The models were created by manually writing the triples. For details, see included README.md</p> <p> </p> <p>[1] J. Granderson, G. Lin, A. Harding, P. Im, Y. Chen, Building fault detection data to aid diagnostic algorithm creation and performance testing, Scientific Data. 7 (2020) 65. https://doi.org/10.1038/s41597-020-0398-6.</p> <p>[2] J.M. House, H. Vaezi-Nejad, J.M. Whitcomb, An expert rule set for fault detection in air-handling units, ASHRAE Transactions. 107 (2001) 858–871.</p>
RDF Subset From PDBj for Evaluating Oxigraph Server
<p>This repository contains the code and data files for the <a href="https://2023.biohackathon.org">DBCLS BioHackathon 2023</a> project "Evaluating Oxigraph Server as a Triple Store for Small and Medium-Sized Datasets".</p> <p>The evaluation focused on the use of [Oxigraph](https://oxigraph.org), a modern graph database implemented in the Rust programming language, as a triple store for handling small to medium-sized (sub-billion) datasets in the context of RDF and SPARQL technologies.</p>
EAGLE Vocabularies RDF data
<p>The data under the EAGLE Vocabularies <a href="https://www.eagle-network.eu/resources/vocabularies/">https://www.eagle-network.eu/resources/vocabularies/</a> .</p> <p>The latest version can always be found in the GitHub repository where these vocabularies are maintained <a href="https://github.com/EAGLE-BPN/epidocupconversion/tree/master/edm%2Bvoc/vocabularies%20testing">https://github.com/EAGLE-BPN/epidocupconversion/tree/master/edm%2Bvoc/vocabularies%20testing</a> .</p> <p>Documentation of the process of initial production of the vocabularies can be found in the project deliverables <a href="https://www.eagle-network.eu/eagle-project/documents-deliverables/">https://www.eagle-network.eu/eagle-project/documents-deliverables/</a> , especially <a href="https://www.eagle-network.eu/wp-content/uploads/2013/06/EAGLE_D2.2.1_Content-harmonisation-guidelines-including-GIS-and-terminologies.pdf">https://www.eagle-network.eu/wp-content/uploads/2013/06/EAGLE_D2.2.1_Content-harmonisation-guidelines-including-GIS-and-terminologies.pdf</a> and <a href="https://www.eagle-network.eu/wp-content/uploads/2013/06/EAGLE_D2.2.2_Content-harmonisation-guidelines-including-GIS-and-terminologies-Second-Release.pdf">https://www.eagle-network.eu/wp-content/uploads/2013/06/EAGLE_D2.2.2_Content-harmonisation-guidelines-including-GIS-and-terminologies-Second-Release.pdf</a></p> <p> </p> <p> </p> <p> </p>
TBFY RDF data from 2019
<p>TheyBuyForYou RDF data from 2019. The file containes the RDF data month by month.</p>
WikiPathways March 2020 Release - GPML and RDF files
<p>Archive of the WikiPathways March 2020 GPML files for <em>Homo sapiens</em> and the GPMLRDF and WPRDF translations. The data is available as CCZero.</p>
Embedding Metadata Shex and Example RDF
<p>Vector Embedding Metadata Example: TransE embeddings for DrugBank computed using TransE.</p>
CoNLL-RDF ontology
<p>The CoNLL-RDF ontology provides machine-readable semantics for an inventory of CoNLL properties (and classes) for a growing collection of about two dozen CoNLL and related formats currently used in language technology.</p>
PROCI RDF
<p>PROCI RDF statements.</p>
Random Wikidata-style RDF + dictionary-based files
Open the record for dataset details and reuse information.
Resource Description Framework (RDF) Modeling of Named Entity Co-occurrences Derived from Biomedical Literature in the PubChemRDF
<p>This is the dataset used for the publication of "Resource Description Framework (RDF) Modeling of Named Entity Co-occurrences Derived from Biomedical Literature in the PubChemRDF". </p>
Semantic Census RDF data
<p>Semantic Census RDF data files</p>
Towards Diversity-Tolerant RDF-Stores
<p>Towards Diversity-Tolerant RDF-Stores</p>
Syntactic Geospatial data generated in RDF format
<p>This dataset represents synthetic generated data from CALLISTO data in RDF form. It contains the equivalent of 2 billion triples in TTL format.</p> <p>Each entity contains:</p> <ul> <li>Crop category: "Grasland" and "Bouwland"</li> <li>Geo information: as Multipolygon in Well Known Text (WKT) format</li> <li>Geometry area</li> <li>Geometry length</li> <li>Object id</li> <li>Parcel</li> <li>Rdf:type owl:NamedIndividual</li> </ul> <p>Send an email to: nagpal@infai.org to have access to the dataset if you need to test GeoSparql query engine on big data.</p>
Dump of RDF dataset used by PO for a Graph Database benchmark, 2022
<p>This dataset represents a newer version of the NQUADS files in RDF from Publication Offices used for benchmarking graph databases. </p> <p> </p>
CLARA Knowledge Graph of licensed educational resources (RDF-star)
<p><span>/!\</span> This deposit is deprecated; a more complete version of the deposit can be found here: <a href="https://zenodo.org/records/8403142">8403142</a>. <span>/!\</span></p> <p><strong>CLARA</strong><br>This deposit is part of the <a href="https://project.inria.fr/clara/">CLARA project</a>. The CLARA project aims to empower teachers in the task of creating new educational resources. And in particular with the task of handling the licenses of reused educational resources.</p> <p>The present deposit contains the RDF files created using an RDF mapping (<a href="https://ceur-ws.org/Vol-2980/paper374.pdf">RML-star</a>) and a mapper (<a href="https://github.com/morph-kgc/morph-kgc">Morph-KGC</a>). The files used as inputs can be found <a href="http://10.5281/zenodo.8107150">here</a>. That pipeline can be found on <a href="https://gitlab.univ-nantes.fr/clara/pipeline">Gitlab</a>. The data used in that pipeline originate from <a href="https://www.x5gon.org/">X5GON</a>, a European project aiming to generate and gather open educational resources.</p> <p><strong>Content</strong><br>The present Knowledge Graph contains information about 45K educational resources and 135K subjects (extracted from DBpedia).<br>That information contains </p> <ul> <li>the author,</li> <li>its title and description</li> <li>the license,</li> <li>a URL to the resource itself,</li> <li>the language of the educational resource,</li> <li>its mimetype,</li> <li>and finally which subject it talks about, and to what extent.</li> </ul> <p><br>That extent is given by two scores, a PageRank score and a cosinus score (that were <strong>not</strong> calculated on our graph but during the X5GON project)</p> <p>A particularity of the Knowledge Graph is its heavy use of RDF reification, across large multi-valued properties.<br>Other versions of the same deposit can be found using different reification models:</p> <ul> <li><a href="https://zenodo.org/record/8108856">Standard reification</a></li> <li><a href="https://zenodo.org/record/8108948">Named graphs</a></li> <li><a href="https://zenodo.org/record/8108963">Singleton properties</a></li> </ul> <p>The Knowledge Graph also contains <a href="https://databus.dbpedia.org/dbpedia/generic/categories">categories</a> originating from DBpedia. They help precise the subjects that are also extracted from DBpedia.</p> <p>The deposit contains five types of files:</p> <ul> <li><strong>Authors_[</strong>X<strong>].nt</strong> - Those contain the authors' nodes, their type, and name.</li> <li><strong>ER_[</strong>X<strong>].ttl</strong> - Those contain the educational resources and their information using RDF-star.</li> <li><strong>categories_skos_[</strong>X<strong>].ttl</strong> - Those contain the hierarchy of DBpedia categories.</li> <li><strong>categories_labels.ttl </strong>- This file contains additional information about the categories.</li> <li><strong>categories_article.ttl</strong> - This file contains the RDF triples that link the DBpedia subjects to the DBpedia categories.</li> </ul>
The Wind in our Sails: BGB DAS RDF dataset
<p>The Wind in our Sails. BGB Dutch Asiatic Shipping dataset converted to RDF. Based on Database Dutch-Asiatic Shipping in the 17th and 18th centuries, 1595-1795 (Huygens ING) http://resources.huygens.knaw.nl/das</p> <p>This deposit contains two parts: DAS and NODAS. For each part this contains the resulting RDF triples in separate zip archives as well as scripts used for conversion and linking.</p>
The Wind in our Sails: DAS RDF dataset
<p>The Wind in our Sails. Dutch Asiatic Shipping dataset converted to RDF. Based on Database Dutch-Asiatic Shipping in the 17th and 18th centuries, 1595-1795 (Huygens ING) http://resources.huygens.knaw.nl/das</p> <p>This deposit contains the resulting RDF triples in separate zip archive as well as scripts used for conversion and linking.</p>
The Wind in our Sails: Places RDF dataset
<p>The Wind in our Sails. Places dataset converted to RDF. </p> <p>This deposit contains the resulting RDF triples in separate zip archives as well as scripts used for conversion and linking.</p>
The Wind in our Sails: VOCOPV RDF dataset
<p>The Wind in our Sails. VOP Opvarenden dataset converted to RDF. Based on VOC Opvarenden as published by the National Archives of the Netherlands https://www.nationaalarchief.nl/onderzoeken/zoekhulpen/voc-opvarenden 1699-1794</p> <p>This deposit contains the resulting RDF triples in separate zip archives as well as scripts used for conversion and linking.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.