Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
100
datasets available to search
ShareScore release 0.7.1
Dataset results
100 results for “RDF”
RDF version of the supplementary data from Shin, Hyun Kil and Seo et al. Meta-analysis of Daphnia magna nanotoxicity experiments in accordance with test guidelines. Environ. Sci.: Nano (2018)
<p>This is an RDF version of the dataset published by Shin, Hyun Kil and Seo et al. as a supplement of the study Meta-analysis of Daphnia magna nanotoxicity experiments in accordance with test guidelines. Environ. Sci.: Nano (2018).</p> <p>The original dataset is available online: <a href="https://ui.staging.kit.cloud.douglasconnect.com/dataexplorer?dataset=ab2bc1ee-99dc-4ddf-b1f9-9fdeb8a0f48c%3A1&q=%7B%7D">https://ui.staging.kit.cloud.douglasconnect.com/dataexplorer?dataset=ab2bc1ee-99dc-4ddf-b1f9-9fdeb8a0f48c%3A1&q=%7B%7D</a></p> <p>The original publication DOI: <a href="http://dx.doi.org/10.1039/C7EN01127J">http://dx.doi.org/10.1039/C7EN01127J</a></p> <p>GitHub repository of the datasets converted to RDF along with RML mappings: <a href="https://github.com/ammar257ammar/RDFied-datasets">https://github.com/ammar257ammar/RDFied-datasets</a></p>
Timestamped RDF-star datasets according to the rdfstar policy
<p>All three datasets were built on the basis of independent snapshots of the <a href="https://aic.ai.wu.ac.at/qadlod/bear.html">BEAR datasets</a>.</p> <p>Each triple was together with a creation and deletion timestamp cast into a <a href="https://w3c.github.io/rdf-star/tests/sparql/syntax/manifest.html#sparql-star-nested-1">nested quoted triples in subject position</a>. </p> <p>The timestamps for both datasets do not reflect the original timestamps of the triples but were for the sake of evaluating the RDF-star policies incremented by 1sec from the initial system timestamp.</p> <p>The main script used to construct these datasets can be found <a href="https://github.com/GreenfishK/starvers_eval/blob/master/scripts_dev/3_construct_datasets/construct_datasets.py">here</a>, which contains the logic on how to assemble these datasets. However, the script was run within a Docker container found on the <a href="https://github.com/GreenfishK/starvers_eval">starvers_eval</a> Github page.</p> <p> </p>
SciQA benchmark: Dataset and RDF dump
<p>SciQA benchmark of questions and queries.</p> <p>The data dump is in NTriples format (RDF NT) taken from the ORKG system on 14.02.2023 at 02:04PM.<br> The dump can be imported into a virtuoso endpoint or any RDF engine so it can be queried.</p> <p>The questions/queries are provided as JSON files, also train and test files are provided for each of the sets. </p> <p><strong>Types</strong> of questions and queries:</p> <ul> <li>Handcrafted set of 100 questions</li> <li>Auto-generated set of 2465 questions</li> </ul> <p><strong>More details</strong> on certain columns:<br> "Classification rationale" It may contain the following values:</p> <ul> <li>Nested facts in the question</li> <li>Sorting, sum, average, minimum, maximum or count calculation required</li> <li>Filter used</li> <li>Mappings of Asking Point in the question to the ORKG ontology</li> </ul> <p><strong>Explanation of Rationale for Non-factoid</strong>:</p> <ul> <li>Nested facts in the question. An entity (e.g., a system or a paper) or predicate is requested that is not explicitly stated in the question text and must be inferred while searching for an answer. </li> <li>Sorting, sum, average, minimum, maximum or count calculation required. To get the answer to the question it is necessary to make an aggregation of the query results. </li> <li>Filter used. To get the answer to the question it is necessary to use filtering of the query results by some conditions.<br> </li> </ul>
CRMSurv Ontology RDF
<p>This ontology extends the CIDOC CRM 6.2 ontology data standard for cultural heritage data in order to provide classes and properties necessary to describe unique aspects of the archaeological survey process. It also makes use of classes and properties of the CIDOC CRM extensions CRMArchaeo 1.4.1 and CRMsci 1.2.3. The ontology is intended to support researchers interested in integrating archaeology survey data using the CIDOC CRM.</p>
RDF Turtle Trip Advisor restaurant NYC test
<p>The RDF Turtle distribution of the Trip Advisor Newyork City restaurants Dataset used in the NeIC spring 2023 FAIR training.</p>
Wikidata RDF Dump of 2020-01-27
<p>This dataset contains the Wikidata RDF full dump of 2020-01-27. The dataset consists in 12 tar files, each one containing at most 100 compressed n-triples files. To get the dataset download the files in a folder and extract all tar files there. The original dump was split to facilitate the parallelization of its processing.</p>
Plazi Treatment RDF Archive
<p>Plazi is an association supporting and promoting the development of persistent and openly accessible digital taxonomic literature. To this end Plazi will:</p> <ul> <li>Maintain a digital taxonomic literature repository to enable archiving of taxonomic treatments.</li> <li>Enhance submitted taxonomic treatments by creating TaxonX and Taxpub XML versions.</li> <li>Participate in the development of new models for publishing taxonomic treatments in order to maximize interoperability with other relevant cyberinfrastructure components (e.g., name servers, biodiversity resources, etc...)</li> <li>Advocate and educate about the vital importance of maintaining free and open access to scientific discourse and data</li> </ul> <p>This publication contains a snapshot of the RDF data associated with over 300,000 taxonomic literature indexed by Plazi accessed via https://github.com/plazi/treatments-rdf/archive/master.zip on 2020-10-01 .</p>
EAGLE Media Wiki RDF data from Wikibase
<p>This is a dump of the triples entered with the Wikibase Extension into the EAGLE Media Wiki for translations.</p> <p>https://wiki.eagle-network.eu/wiki/Main_Page </p> <p>The same data is accessible via the Mediawiki API. </p> <p>It is part of the EAGLE project https://www.eagle-network.eu/.</p> <p> </p> <p>Contributors of the translations are in the data.</p>
tdwg-challenge-rdf: v1.0
<p>First release</p>
Itag2.3 Tomato Genome Annotation, RDF graph
<p>Annotation of the tomato genome, ITAG2.3 (ftp://ftp.sgn.cornell.edu/genomes/Solanum_lycopersicum/annotation/ITAG2.3_release/). GFF file's were converted into a RDF graph.</p>
Beta version of OpenResearch.org RDF dump
<p>This RDF dump containg the data about scientific events that are currently stored as wiki pages in OpenResearch.org. </p>
NorwegianSoE RDF dump
<p>This is a RDF dump of the Norwegian State of Estate Report dataset. The dataset is produced by integrating cross-domain government datasets including data from sources such as the Norwegian business entity register, cadastral system, building accessibility register and the previous SoE report.</p>
biotea i-n dataset of rdf for pmc
<p>i-n dataset of rdf for pmc</p>
biotea, RDf for pubmed central, c-h dataset
<p>C-H dataset of the RDF for pubmed central </p>
RE-DWELL Knowledge base dataset (RDF)
<p>This RDF dataset represents a collection of concepts and case studies focused on urban and housing projects, capturing a wide range of details such as project descriptions, types, locations, timeframes, construction systems, and associated Sustainable Development Goals (SDGs).</p> <p>It includes metadata about the authors of each study, related concepts, and the geographic contexts of the projects. </p> <p>The dataset is designed to support analysis and exploration of trends, relationships, and insights in urban development, housing strategies, and their alignment with broader social and environmental goals. </p> <p>It provides a rich foundation for research in architecture, urban planning, and sustainability studies.</p> <p> </p> <p> </p> <p> </p>
WikiPathways Sept 2021 Release - RDF data
<p>Archive of the WikiPathways September 2021 RDF data for <em>Homo sapiens</em> as GPMLRDF and WPRDF (.ttl format). The data is licenced under the <a href="https://creativecommons.org/share-your-work/public-domain/cc0/">CCZero waiver</a>.</p>
Publications Office SPARQL queries for RDF benchmark
<p>This dataset contains the queries used for the benchmarking of the RDF stores for the EU Publications Office (PO).</p> <p>The queries are divided in three categories according to the types of queries. The initial queries from PO where validated using Jena tool to make them compliant with any SPARQL endpoint. </p>
RDF Dataset for article: A confidence predictor for logD using conformal regression and a support-vector machine
<p>RDF dataset described in article: "A confidence predictor for logD using conformal regression and a support-vector machine" (Manuscript in preparation).</p> <p>The dataset contains conformal logD values at 90% confidence level, computed for 91M compounds from PubChem, in RDF format.</p> <p>The .hdt.gz version contains the dataset in RDF HDT format (http://www.rdfhdt.org/), compressed with tar and gzip. The archive contains both the .hdt file, and an index file, generated by the hdtSearch C++ tool.</p> <p>The .ttl.gz file is a gzipped file in RDF Turtle format (https://www.w3.org/TR/turtle/).</p>
Benchmark results for RDF load time evaluation (RiverBench)
<p>Benchmark results of an RDF load time evaluation.</p> <p>Benchmark code: <a href="https://github.com/Ostrzyciel/rdf4led-riverbench">https://github.com/Ostrzyciel/rdf4led-riverbench</a></p> <p>Input data: <a href="https://doi.org/10.5281/zenodo.12073223">https://doi.org/10.5281/zenodo.12073223</a></p> <p><strong>Please see the README.md file for a detailed description of the data format.</strong></p>
RDF triples of raw and trajectory synopses over AIS kinematic messages in Brest
<p>The following list of data sets is derived from the final data set of trajectory synopses available at https://zenodo.org/record/2563256 . We have converted the original data into a) ESRI shapefiles and b) using RDF-Gen (https://zenodo.org/record/2556747) into RDF triples w.r.t. the datAcron ontology.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.