Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

209

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

209 results for “KG”

Learn how ShareScore rates datasets ↗
zenodo44/100

Silt content in % (kg / kg) at 6 standard depths (0, 10, 30, 60, 100 and 200 cm) at 250 m resolution

<p>Silt content in % (kg / kg) at 6 standard depths (0, 10, 30, 60, 100 and 200 cm) at 250 m resolution.&nbsp;Based on machine learning predictions from global compilation of soil profiles and samples. Processing steps are described in detail <strong><a href="https://gitlab.com/openlandmap/global-layers/tree/master/soil">here</a></strong>. Antarctica is not included.</p> <p>To access and visualize maps use:&nbsp;&nbsp;<a href="http://www.openlandmap.org/">OpenLandMap.org</a></p> <p>If you discover a bug, artifact or inconsistency in the maps, or if you have a question please use some of the following channels:</p> <ul> <li>Technical issues and questions about the code:&nbsp;<a href="https://gitlab.com/openlandmap/global-layers/issues">https://gitlab.com/openlandmap/global-layers/issues</a>&nbsp;</li> <li>General questions and comments:&nbsp;<a href="https://disqus.com/home/forums/landgis/">https://disqus.com/home/forums/landgis/</a></li> </ul> <p>&nbsp;</p> <p>All files internally compressed using &quot;COMPRESS=DEFLATE&quot; creation&nbsp;option in GDAL. File naming convention:</p> <ul> <li>sol = theme: soil,</li> <li>silt.wfraction = variable: silt weight fraction,</li> <li>usda.3a1a1a = determination method: laboratory method code,</li> <li>m = mean value,</li> <li>250m = spatial resolution / block support: 250 m,</li> <li>b10..10cm = vertical reference: 10 cm depth below surface,</li> <li>1950..2017 = time reference: period 1950-2017,</li> <li>v0.2 = version number: 0.2,</li> </ul>

opencc-by-nc-sa-4.0Dec 2018View details →
zenodo44/100

The URW-KG: a Resource for Tackling the Under-Representation of non-Western Writers

<p>Digital media have enabled the access to an unprecedented literary knowledge. Authors, readers, and scholars are now able to discover and share an increasing amount of information about books and their authors. Notwithstanding, digital archives are still unbalanced: writers from non-Western countries are less represented, and such a condition leads to the perpetration of old forms of discrimination. In this paper, we present the Under-Represented Writers Knowledge Graph (URW-KG), a resource designed to explore and possibly amend this lack of representation by gathering and mapping information about works and authors from Wikidata and three other sources: Open Library, Goodreads, and Google Books. The experiments based on KG embeddings showed that the integrated information encoded in the graph allows scholars and users to be more easily exposed to non-Western literary works and authors with respect to Wikidata alone. This opens to the development of fairer and effective tools for author discovery and exploration.<br> &nbsp;</p>

opencc-by-4.0Dec 2022View details →
zenodo44/100

KG for heart failure gene expression data

<p>Pre processed&nbsp;gene expression data for&nbsp;different heart failure. Includes count table,&nbsp; gene patiens metadata, gene lenght</p>

opencc-by-4.0Mar 2023View details →
zenodo40/100

PheKnowLator Human Disease KG Benchmarks: Instance-Inverse Relations-OWLNETS (v2.1.0 - May 2021)

<p><strong>PKT Human Disease Knowledge Graph Benchmark Builds&nbsp;(v2.1.0)</strong></p><p><strong>Build Type:&nbsp;</strong><i>Instance-Inverse Relations-OWLNETS</i></p><p><strong>Build Date: </strong>May&nbsp;01, 2021</p><p>&nbsp;</p><h3><strong>Important Build Information</strong></h3><p>The benchmarks were originally built and stored using Google Cloud Platform (GCP) resources. For details and a complete description of this process, can be found on GitHub (<a href="https://github.com/callahantiff/PheKnowLator/tree/master/builds#readme">here</a>). Note that we have developed an archive for the builds on Zenodo. While the original GCP resources contained all associated files, due to the file size upload limits associated with each archive, we have limited the uploaded files to the KGs, associated metadata, and log files. The list of resources, including their URLs, and date of download, can all be found in the associated logs.</p><p>Details on each of the files generated by the build process can be found in the file associated with this directory (<a href="https://zenodo.org/records/10065431/files/PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx?download=1">PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx</a>).</p><p>&nbsp;</p><p>🚨&nbsp;<strong>AVAILABLE FILES&nbsp;</strong>🚨&nbsp;</p><ul><li>Available KG benchmark files are zipped and listed below.</li><li>For additional details on what each file contains, please see the associated Wiki page&nbsp;👉&nbsp;<a href="https://github.com/callahantiff/PheKnowLator/wiki/May-01%2C-2021">here</a>.</li></ul>

opencc-by-4.0Apr 2021View details →
zenodo40/100

PheKnowLator Human Disease KG Benchmarks: Instance-Standard Relations-OWLNETS (v2.1.0 - May 2021)

<p><strong>PKT Human Disease Knowledge Graph Benchmark Builds&nbsp;(v2.1.0)</strong></p><p><strong>Build Type:&nbsp;</strong><i>Instance-Standard Relations-OWLNETS</i></p><p><strong>Build Date: </strong>May&nbsp;01, 2021</p><p>&nbsp;</p><h3><strong>Important Build Information</strong></h3><p>The benchmarks were originally built and stored using Google Cloud Platform (GCP) resources. For details and a complete description of this process, can be found on GitHub (<a href="https://github.com/callahantiff/PheKnowLator/tree/master/builds#readme">here</a>). Note that we have developed an archive for the builds on Zenodo. While the original GCP resources contained all associated files, due to the file size upload limits associated with each archive, we have limited the uploaded files to the KGs, associated metadata, and log files. The list of resources, including their URLs, and date of download, can all be found in the associated logs.</p><p>Details on each of the files generated by the build process can be found in the file associated with this directory (<a href="https://zenodo.org/records/10065431/files/PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx?download=1">PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx</a>).</p><p>&nbsp;</p><p>🚨&nbsp;<strong>AVAILABLE FILES&nbsp;</strong>🚨&nbsp;</p><ul><li>Available KG benchmark files are zipped and listed below.</li><li>For additional details on what each file contains, please see the associated Wiki page&nbsp;👉&nbsp;<a href="https://github.com/callahantiff/PheKnowLator/wiki/May-01%2C-2021">here</a>.</li></ul>

opencc-by-4.0Apr 2021View details →
zenodo40/100

PheKnowLator Human Disease KG Benchmarks: Instance-Standard Relations-OWL (v2.1.0 - May 2021)

<p><strong>PKT Human Disease Knowledge Graph Benchmark Builds&nbsp;(v2.1.0)</strong></p><p><strong>Build Type:&nbsp;</strong><i>Instance-Standard Relations-OWL</i></p><p><strong>Build Date: </strong>May&nbsp;01, 2021</p><p>&nbsp;</p><h3><strong>Important Build Information</strong></h3><p>The benchmarks were originally built and stored using Google Cloud Platform (GCP) resources. For details and a complete description of this process, can be found on GitHub (<a href="https://github.com/callahantiff/PheKnowLator/tree/master/builds#readme">here</a>). Note that we have developed an archive for the builds on Zenodo. While the original GCP resources contained all associated files, due to the file size upload limits associated with each archive, we have limited the uploaded files to the KGs, associated metadata, and log files. The list of resources, including their URLs, and date of download, can all be found in the associated logs.</p><p>Details on each of the files generated by the build process can be found in the file associated with this directory (<a href="https://zenodo.org/records/10065431/files/PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx?download=1">PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx</a>).</p><p>&nbsp;</p><p>🚨&nbsp;<strong>AVAILABLE FILES&nbsp;</strong>🚨&nbsp;</p><ul><li>Available KG benchmark files are zipped and listed below.</li><li>For additional details on what each file contains, please see the associated Wiki page&nbsp;👉&nbsp;<a href="https://github.com/callahantiff/PheKnowLator/wiki/May-01%2C-2021">here</a>.</li></ul>

opencc-by-4.0Apr 2021View details →
zenodo40/100

PheKnowLator Human Disease KG Benchmarks: Instance-Inverse Relations-OWL (v2.1.0 - May 2021)

<p><strong>PKT Human Disease Knowledge Graph Benchmark Builds&nbsp;(v2.1.0)</strong></p><p><strong>Build Type:&nbsp;</strong><i>Instance-Inverse Relations-OWL</i></p><p><strong>Build Date: </strong>May&nbsp;01, 2021</p><p>&nbsp;</p><h3><strong>Important Build Information</strong></h3><p>The benchmarks were originally built and stored using Google Cloud Platform (GCP) resources. For details and a complete description of this process, can be found on GitHub (<a href="https://github.com/callahantiff/PheKnowLator/tree/master/builds#readme">here</a>). Note that we have developed an archive for the builds on Zenodo. While the original GCP resources contained all associated files, due to the file size upload limits associated with each archive, we have limited the uploaded files to the KGs, associated metadata, and log files. The list of resources, including their URLs, and date of download, can all be found in the associated logs.</p><p>Details on each of the files generated by the build process can be found in the file associated with this directory (<a href="https://zenodo.org/records/10065431/files/PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx?download=1">PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx</a>).</p><p>&nbsp;</p><p>🚨&nbsp;<strong>AVAILABLE FILES&nbsp;</strong>🚨&nbsp;</p><ul><li>Available KG benchmark files are zipped and listed below.</li><li>For additional details on what each file contains, please see the associated Wiki page&nbsp;👉&nbsp;<a href="https://github.com/callahantiff/PheKnowLator/wiki/May-01%2C-2021">here</a>.</li></ul>

opencc-by-4.0Apr 2021View details →
zenodo40/100

PheKnowLator Human Disease KG Benchmarks: Class-Standard Relations-OWLNETS (v2.1.0 - May 2021)

<p><strong>PKT Human Disease Knowledge Graph Benchmark Builds&nbsp;(v2.1.0)</strong></p><p><strong>Build Type:&nbsp;</strong><i>Class-Standard Relations-OWLNETS</i></p><p><strong>Build Date: </strong>May&nbsp;01, 2021</p><p>&nbsp;</p><h3><strong>Important Build Information</strong></h3><p>The benchmarks were originally built and stored using Google Cloud Platform (GCP) resources. For details and a complete description of this process, can be found on GitHub (<a href="https://github.com/callahantiff/PheKnowLator/tree/master/builds#readme">here</a>). Note that we have developed an archive for the builds on Zenodo. While the original GCP resources contained all associated files, due to the file size upload limits associated with each archive, we have limited the uploaded files to the KGs, associated metadata, and log files. The list of resources, including their URLs, and date of download, can all be found in the associated logs.</p><p>Details on each of the files generated by the build process can be found in the file associated with this directory (<a href="https://zenodo.org/records/10065431/files/PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx?download=1">PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx</a>).</p><p>&nbsp;</p><p>🚨&nbsp;<strong>AVAILABLE FILES&nbsp;</strong>🚨&nbsp;</p><ul><li>Available KG benchmark files are zipped and listed below.</li><li>For additional details on what each file contains, please see the associated Wiki page&nbsp;👉&nbsp;<a href="https://github.com/callahantiff/PheKnowLator/wiki/May-01%2C-2021">here</a>.</li></ul>

opencc-by-4.0Apr 2021View details →
zenodo40/100

PheKnowLator Human Disease KG Benchmarks: Instance-Standard Relations-OWL (v2.0.0 - February 2021)

<p><strong>PKT Human Disease Knowledge Graph Benchmark Builds&nbsp;(v2.0.0)</strong></p><p><strong>Build Type:&nbsp;</strong><i>Instance-Standard Relations-OWL</i></p><p><strong>Build Date: </strong>February 11, 2021</p><p>&nbsp;</p><h3><strong>Important Build Information</strong></h3><p>The benchmarks were originally built and stored using Google Cloud Platform (GCP) resources. For details and a complete description of this process, can be found on GitHub (<a href="https://github.com/callahantiff/PheKnowLator/tree/master/builds#readme">here</a>). Note that we have developed an archive for the builds on Zenodo. While the original GCP resources contained all associated files, due to the file size upload limits associated with each archive, we have limited the uploaded files to the KGs, associated metadata, and log files. The list of resources, including their URLs, and date of download, can all be found in the associated logs.</p><p>Details on each of the files generated by the build process can be found in the file associated with this directory (<a href="https://zenodo.org/records/10065431/files/PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx?download=1">PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx</a>).</p><p>&nbsp;</p><p>🚨&nbsp;<strong>AVAILABLE FILES&nbsp;</strong>🚨&nbsp;</p><ul><li>Available KG benchmark files are zipped and listed below.</li><li>For additional details on what each file contains, please see the associated Wiki page&nbsp;👉&nbsp;<a href="https://github.com/callahantiff/PheKnowLator/wiki/February-11%2C-2021">here</a>.</li></ul>

opencc-by-4.0Jan 2021View details →
zenodo40/100

PheKnowLator Human Disease KG Benchmarks: Instance-Standard Relations-OWLNETS (v2.0.0 - February 2021)

<p><strong>PKT Human Disease Knowledge Graph Benchmark Builds&nbsp;(v2.0.0)</strong></p><p><strong>Build Type:&nbsp;</strong><i>Class-Standard Relations-OWLNETS</i></p><p><strong>Build Date: </strong>February 11, 2021</p><p>&nbsp;</p><h3><strong>Important Build Information</strong></h3><p>The benchmarks were originally built and stored using Google Cloud Platform (GCP) resources. For details and a complete description of this process, can be found on GitHub (<a href="https://github.com/callahantiff/PheKnowLator/tree/master/builds#readme">here</a>). Note that we have developed an archive for the builds on Zenodo. While the original GCP resources contained all associated files, due to the file size upload limits associated with each archive, we have limited the uploaded files to the KGs, associated metadata, and log files. The list of resources, including their URLs, and date of download, can all be found in the associated logs.</p><p>Details on each of the files generated by the build process can be found in the file associated with this directory (<a href="https://zenodo.org/records/10065431/files/PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx?download=1">PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx</a>).</p><p>&nbsp;</p><p>🚨&nbsp;<strong>AVAILABLE FILES&nbsp;</strong>🚨&nbsp;</p><ul><li>Available KG benchmark files are zipped and listed below.</li><li>For additional details on what each file contains, please see the associated Wiki page&nbsp;👉&nbsp;<a href="https://github.com/callahantiff/PheKnowLator/wiki/February-11%2C-2021">here</a>.</li></ul>

opencc-by-4.0Jan 2021View details →
zenodo40/100

PheKnowLator Human Disease KG Benchmarks: Instance-Standard Relations-OWL (v2.0.0 - January 2021)

<p><strong>PKT Human Disease Knowledge Graph Benchmark Builds&nbsp;(v2.0.0)</strong></p><p><strong>Build Type:&nbsp;</strong><i>Instance-Standard&nbsp;Relations-OWL</i></p><p><strong>Build Date: </strong>January 25, 2021</p><p>&nbsp;</p><h3><strong>Important Build Information</strong></h3><p>The benchmarks were originally built and stored using Google Cloud Platform (GCP) resources. For details and a complete description of this process, can be found on GitHub (<a href="https://github.com/callahantiff/PheKnowLator/tree/master/builds#readme">here</a>). Note that we have developed an archive for the builds on Zenodo. While the original GCP resources contained all associated files, due to the file size upload limits associated with each archive, we have limited the uploaded files to the KGs, associated metadata, and log files. The list of resources, including their URLs, and date of download, can all be found in the associated logs.</p><p>Details on each of the files generated by the build process can be found in the file associated with this directory (<a href="https://zenodo.org/records/10065431/files/PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx?download=1">PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx</a>).</p><p>&nbsp;</p><p>🚨&nbsp;<strong>AVAILABLE FILES&nbsp;</strong>🚨&nbsp;</p><ul><li>Available KG benchmark files are zipped and listed below.</li><li>For additional details on what each file contains, please see the associated Wiki page&nbsp;👉&nbsp;<a href="https://github.com/callahantiff/PheKnowLator/wiki/January-25%2C-2021">here</a>.</li></ul>

opencc-by-4.0Jan 2021View details →
zenodo40/100

PheKnowLator Human Disease KG Benchmarks: Class-Standard Relations-OWLNETS (v2.0.0 - February 2021)

<p><strong>PKT Human Disease Knowledge Graph Benchmark Builds&nbsp;(v2.0.0)</strong></p><p><strong>Build Type:&nbsp;</strong><i>Class-Standard Relations-OWLNETS</i></p><p><strong>Build Date: </strong>February 11, 2021</p><p>&nbsp;</p><h3><strong>Important Build Information</strong></h3><p>The benchmarks were originally built and stored using Google Cloud Platform (GCP) resources. For details and a complete description of this process, can be found on GitHub (<a href="https://github.com/callahantiff/PheKnowLator/tree/master/builds#readme">here</a>). Note that we have developed an archive for the builds on Zenodo. While the original GCP resources contained all associated files, due to the file size upload limits associated with each archive, we have limited the uploaded files to the KGs, associated metadata, and log files. The list of resources, including their URLs, and date of download, can all be found in the associated logs.</p><p>Details on each of the files generated by the build process can be found in the file associated with this directory (<a href="https://zenodo.org/records/10065431/files/PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx?download=1">PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx</a>).</p><p>&nbsp;</p><p>🚨&nbsp;<strong>AVAILABLE FILES&nbsp;</strong>🚨&nbsp;</p><ul><li>Available KG benchmark files are zipped and listed below.</li><li>For additional details on what each file contains, please see the associated Wiki page&nbsp;👉&nbsp;<a href="https://github.com/callahantiff/PheKnowLator/wiki/February-11%2C-2021">here</a>.</li></ul>

opencc-by-4.0Jan 2021View details →
zenodo40/100

PheKnowLator Human Disease KG Benchmarks: Instance-Inverse Relations-OWL (v2.0.0 - February 2021)

<p><strong>PKT Human Disease Knowledge Graph Benchmark Builds&nbsp;(v2.0.0)</strong></p><p><strong>Build Type:&nbsp;</strong><i>Instance-Inverse Relations-OWL</i></p><p><strong>Build Date: </strong>February 11, 2021</p><p>&nbsp;</p><h3><strong>Important Build Information</strong></h3><p>The benchmarks were originally built and stored using Google Cloud Platform (GCP) resources. For details and a complete description of this process, can be found on GitHub (<a href="https://github.com/callahantiff/PheKnowLator/tree/master/builds#readme">here</a>). Note that we have developed an archive for the builds on Zenodo. While the original GCP resources contained all associated files, due to the file size upload limits associated with each archive, we have limited the uploaded files to the KGs, associated metadata, and log files. The list of resources, including their URLs, and date of download, can all be found in the associated logs.</p><p>Details on each of the files generated by the build process can be found in the file associated with this directory (<a href="https://zenodo.org/records/10065431/files/PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx?download=1">PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx</a>).</p><p>&nbsp;</p><p>🚨&nbsp;<strong>AVAILABLE FILES&nbsp;</strong>🚨&nbsp;</p><ul><li>Available KG benchmark files are zipped and listed below.</li><li>For additional details on what each file contains, please see the associated Wiki page&nbsp;👉&nbsp;<a href="https://github.com/callahantiff/PheKnowLator/wiki/February-11%2C-2021">here</a>.</li></ul>

opencc-by-4.0Jan 2012View details →
zenodo40/100

PheKnowLator Human Disease KG Benchmarks: Class-Inverse Relations-OWL (v2.0.0 - February 2021)

<p><strong>PKT Human Disease Knowledge Graph Benchmark Builds&nbsp;(v2.0.0)</strong></p><p><strong>Build Type:&nbsp;</strong><i>Class-Inverse Relations-OWL</i></p><p><strong>Build Date: </strong>February 11, 2021</p><p>&nbsp;</p><h3><strong>Important Build Information</strong></h3><p>The benchmarks were originally built and stored using Google Cloud Platform (GCP) resources. For details and a complete description of this process, can be found on GitHub (<a href="https://github.com/callahantiff/PheKnowLator/tree/master/builds#readme">here</a>). Note that we have developed an archive for the builds on Zenodo. While the original GCP resources contained all associated files, due to the file size upload limits associated with each archive, we have limited the uploaded files to the KGs, associated metadata, and log files. The list of resources, including their URLs, and date of download, can all be found in the associated logs.</p><p>Details on each of the files generated by the build process can be found in the file associated with this directory (<a href="https://zenodo.org/records/10065431/files/PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx?download=1">PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx</a>).</p><p>&nbsp;</p><p>🚨&nbsp;<strong>AVAILABLE FILES&nbsp;</strong>🚨&nbsp;</p><ul><li>Available KG benchmark files are zipped and listed below.</li><li>For additional details on what each file contains, please see the associated Wiki page&nbsp;👉&nbsp;<a href="https://github.com/callahantiff/PheKnowLator/wiki/February-11%2C-2021">here</a>.</li></ul>

opencc-by-4.0Jan 2021View details →
zenodo40/100

PheKnowLator Human Disease KG Benchmarks: Class-Standard Relations-OWL (v2.0.0 - February 2021)

<p><strong>PKT Human Disease Knowledge Graph Benchmark Builds&nbsp;(v2.0.0)</strong></p><p><strong>Build Type:&nbsp;</strong><i>Class-Standard Relations-OWL</i></p><p><strong>Build Date: </strong>February 11, 2021</p><p>&nbsp;</p><h3><strong>Important Build Information</strong></h3><p>The benchmarks were originally built and stored using Google Cloud Platform (GCP) resources. For details and a complete description of this process, can be found on GitHub (<a href="https://github.com/callahantiff/PheKnowLator/tree/master/builds#readme">here</a>). Note that we have developed an archive for the builds on Zenodo. While the original GCP resources contained all associated files, due to the file size upload limits associated with each archive, we have limited the uploaded files to the KGs, associated metadata, and log files. The list of resources, including their URLs, and date of download, can all be found in the associated logs.</p><p>Details on each of the files generated by the build process can be found in the file associated with this directory (<a href="https://zenodo.org/records/10065431/files/PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx?download=1">PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx</a>).</p><p>&nbsp;</p><p>🚨&nbsp;<strong>AVAILABLE FILES&nbsp;</strong>🚨&nbsp;</p><ul><li>Available KG benchmark files are zipped and listed below.</li><li>For additional details on what each file contains, please see the associated Wiki page&nbsp;👉&nbsp;<a href="https://github.com/callahantiff/PheKnowLator/wiki/February-11%2C-2021">here</a>.</li></ul>

opencc-by-4.0Jan 2021View details →
zenodo40/100

PheKnowLator Human Disease KG Benchmarks: Class-Inverse Relations-OWLNETS (v2.0.0 - February 2021)

<p><strong>PKT Human Disease Knowledge Graph Benchmark Builds&nbsp;(v2.0.0)</strong></p><p><strong>Build Type:&nbsp;</strong><i>Class-InverseRelations-OWLNETS</i></p><p><strong>Build Date: </strong>February 11, 2021</p><p>&nbsp;</p><h3><strong>Important Build Information</strong></h3><p>The benchmarks were originally built and stored using Google Cloud Platform (GCP) resources. For details and a complete description of this process, can be found on GitHub (<a href="https://github.com/callahantiff/PheKnowLator/tree/master/builds#readme">here</a>). Note that we have developed an archive for the builds on Zenodo. While the original GCP resources contained all associated files, due to the file size upload limits associated with each archive, we have limited the uploaded files to the KGs, associated metadata, and log files. The list of resources, including their URLs, and date of download, can all be found in the associated logs.</p><p>Details on each of the files generated by the build process can be found in the file associated with this directory (<a href="https://zenodo.org/records/10065431/files/PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx?download=1">PheKnowLator_HumanDiseaseKG_Output_FileInformation.xlsx</a>).</p><p>&nbsp;</p><p>🚨&nbsp;<strong>AVAILABLE FILES&nbsp;</strong>🚨&nbsp;</p><ul><li>Available KG benchmark files are zipped and listed below.</li><li>For additional details on what each file contains, please see the associated Wiki page&nbsp;👉&nbsp;<a href="https://github.com/callahantiff/PheKnowLator/wiki/February-11%2C-2021">here</a>.</li></ul>

opencc-by-4.0Jan 2021View details →
zenodo40/100

EB-KG: Knowledge Graph of the first 8 eiditions Encyclopaedia Brittanica (1768-1860)

<p>This Knowlege Graph represents the information of the first eight editions of Encyclopaedia Brittanica (years: 1768 to 1860) in RDF (ttl format).</p> <p>The raw dataset is provided by the NLS in this <a href="https://data.nls.uk/data/digitised-collections/encyclopaedia-britannica/">link</a> , and it comprises of eight editions and a total of 195 volumes with a total size of 44GB. It uses two XMLs schemas: METS&nbsp; for descriptive, structural, technical and administrative metadata (Title, Author, Publisher, etc); and ALTO&nbsp; for encoding the OCR text of a page.</p> <p>In this work, we have extracted the information from METS and ALTO XMLS using <a href="https://github.com/francesNLP/defoe">defoe</a> tool and developed <a href="https://github.com/francesNLP/defoe/tree/master/defoe/nlsArticles/queries">novel information extraction heuristics</a>. With the extracted information, we created the EB-KG Knowlege Graph, which&nbsp; uses the <a href="https://francesnlp.github.io/EB-ontology/doc/index-en.html">EB Ontolgy</a>, to represent such information. Furthermore, during the information extraction phase, we have employed several techniques to mitigate two common OCR errors: long-S and the line-break hyphenation.</p> <p>The EB-KG contains 1,638,239 RDF triples. It has information from 8 editions. Each edition can have several Volumes, references to Books, Supplements; it also has an Editor and a Publisher, which can be a Person or an Organization. A Volume has several Pages, which can contain several Terms. And a Term can be either a Topic (a term described across several pages, often combining text, pictures, and tables.) or an Article (a description of the term in one- or two-paragraph long text (similar to an entry in a dictionary)). The data model of the EB-KG can be found <a href="https://francesnlp.github.io/EB-ontology/doc/dataModel.png">here</a>.</p> <p>The original ALTO files do not indicate the start and end of each EB term, the first part of our work involved the<br> automated extraction of all terms (along with their metadata) across editions, so they can be analysed independently without the surrounding text.</p> <p>&nbsp;</p> <p>&nbsp;</p> <p>&nbsp;</p>

opencc-by-4.0Jun 2022View details →
zenodo40/100

GazetteersScotland-KG: A Knowlege Graph for representing the Gazetteers of Scotland (1803-1901)

<p>This Knowlege Graph represents the information of the &quot;Gazeteers of Scotland<strong>&quot;</strong> (years: 1803 - 1901) collection in RDF (ttl format). This collection comprises twenty volumes of the most popular descriptive gazetteers of Scotland in the 19th century. Principal places in Scotland, including towns, counties, castles, glens, antiquities and parishes, are listed alphabetically. Each entry includes detailed historical and geographical information about each place.&nbsp; The raw dataset is provided by the NLS in this <a href="https://data.nls.uk/data/digitised-collections/gazetteers-of-scotland/">link</a>. As&nbsp; other NLS data collections, they are originally provided using two XMLs schemas: METS&nbsp; for descriptive, structural, technical and administrative metadata (Title, Author, Publisher, etc); and ALTO&nbsp; for encoding the OCR text of a page.</p> <p>In this work, we have extracted the information from METS and ALTO XMLS using <a href="https://github.com/francesNLP/defoe">defoe</a> tool and developed a <a href="https://github.com/francesNLP/defoe/blob/master/defoe/nls/queries/write_metadata_pages_yml.py">new information extraction defoe query</a> , and created a new Knowlege Graph called GazetteersScotland-KG.&nbsp; The GazetteersScotland-KG uses the <a href="https://francesnlp.github.io/NLS-ontology/doc/index-en.html">NLS Ontology </a>to represent the information extracted. Furthermore, during the information extraction phase, we have employed several techniques to mitigate two common OCR errors: long-S and the line-break hyphenation.</p> <p>The GazetteersScotland-KG contains&nbsp;354,998 RDF triples. It has information from 12 series and 20 volumes: Each serie can have several Volumes. Each serie has an Editorm Publisher, mmsid, Shelf-Locator, publication year, etc.&nbsp; A Volume has several Pages,&nbsp; with text in them. The data model of the GazetteersScotland-KG can be found <a href="https://francesnlp.github.io/NLS-ontology/doc/dataModel.png">here</a>.</p> <pre> &nbsp;</pre>

opencc-by-4.0Jun 2022View details →
zenodo40/100

LadiesDebating-KG: A Knowlege Graph for representing the "Edinburgh Ladies' Debating Society Digital Collection" (1865 - 1880)

<p>This Knowlege Graph represents the information of the &quot;Edinburgh Ladies&rsquo; Debating Society<strong>&quot;</strong> (years: 1865 - 1880) collection in RDF (ttl format). This collection consists of the complete runs of two Edinburgh journals, <strong>&lsquo;The Attempt&rsquo; (10 volumes, 1865-74)</strong> and its successor &lsquo;<strong>The Ladies&rsquo; Edinburgh Magazine&rsquo; (6 volumes, 1875-80)</strong>. These publications were produced by a leading Edinburgh women&rsquo;s club, known during the period as the Edinburgh Essay Society or the Ladies&rsquo; Edinburgh Essay Society, but subsequently as the Ladies&rsquo; Edinburgh Debating Society. The Society existed from 1865 to 1935.&nbsp; The raw dataset is provided by the NLS in this <a href="https://data.nls.uk/data/digitised-collections/edinburgh-ladies-debating-society/">link</a>. As&nbsp; other NLS data collections, they are originally provided using two XMLs schemas: METS&nbsp; for descriptive, structural, technical and administrative metadata (Title, Author, Publisher, etc); and ALTO&nbsp; for encoding the OCR text of a page.</p> <p>In this work, we have extracted the information from METS and ALTO XMLS using <a href="https://github.com/francesNLP/defoe">defoe</a> tool and developed a <a href="https://github.com/francesNLP/defoe/blob/master/defoe/nls/queries/write_metadata_pages_yml.py">new information extraction defoe query</a> , and created a new Knowlege Graph called LadiesDebating-KG.&nbsp; The LadiesDebating-KG uses the <a href="https://francesnlp.github.io/NLS-ontology/doc/index-en.html">NLS Ontology </a>to represent the information extracted. Furthermore, during the information extraction phase, we have employed several techniques to mitigate two common OCR errors: long-S and the line-break hyphenation.</p> <p>The LadiesDebating-KG contains 38,279 RDF triples. It has information from 2 series and 16 volumes: <strong>&#39;The attempt&#39; </strong>serie has 10 volumes and&nbsp; <strong>&#39;The Ladies&#39; </strong>serie<strong> </strong>has 6 volumes . Each serie has an Editor,&nbsp;mmsid, Shelf-Locator, publication year, etc.&nbsp; A Volume has several Pages,&nbsp; with text in them. The data model of the LadiesDebating-KG can be found <a href="https://francesnlp.github.io/NLS-ontology/doc/dataModel.png">here</a>.</p> <pre>&nbsp;</pre>

opencc-by-4.0Jun 2022View details →
zenodo40/100

ChapbooksScotland-KG: A Knowlege Graph for representing the "Chapbooks Printed In Scotland" (1671 - 1893)

<p>This Knowlege Graph represents the information of the &quot;<strong>Chapbooks Printed In Scotland&quot;</strong> (years: 1671 - 1893) collection in RDF (ttl format). This dataset comprises more than 3,000 chapbooks printed in Scotland from the 17th to 19th century. They form part of the Lauriston Castle Collection, which was bequeathed to the Library in 1926. It includes some 500 chapbook volumes containing around 5,500 individual items, more than half of which were printed in Scotland.&nbsp; The raw dataset is provided by the NLS in this <a href="https://data.nls.uk/data/digitised-collections/chapbooks-printed-in-scotland/">link</a>. As&nbsp; other NLS data collections, they are originally provided using two XMLs schemas: METS&nbsp; for descriptive, structural, technical and administrative metadata (Title, Author, Publisher, etc); and ALTO&nbsp; for encoding the OCR text of a page.</p> <p>In this work, we have extracted the information from METS and ALTO XMLS using <a href="https://github.com/francesNLP/defoe">defoe</a> tool and developed a <a href="https://github.com/francesNLP/defoe/blob/master/defoe/nls/queries/write_metadata_pages_yml.py">new information extraction defoe query</a> , and created a new Knowlege Graph called ChapbooksScotland-KG.&nbsp; The ChapbooksScotland-KG uses the <a href="https://francesnlp.github.io/NLS-ontology/doc/index-en.html">NLS Ontology </a>to represent the information extracted. Furthermore, during the information extraction phase, we have employed several techniques to mitigate two common OCR errors: long-S and the line-break hyphenation.</p> <p>The ChapbooksScotland-KG contains 352,270 RDF triples. It has information from 2728 series and 3080 volumes. Each serie can have several Volumes, Suplements, references to Books; it also has an Editor and a Publisher, which can be a Person or an Organization. A Volume has several Pages,&nbsp; with text in them. The data model of the ChapbooksScotland-KG can be found <a href="https://francesnlp.github.io/NLS-ontology/doc/dataModel.png">here</a>.</p>

opencc-by-4.0Jun 2022View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record