Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
9
datasets available to search
ShareScore release 0.9.0
Dataset results
9 results for “cidoc crm”
CIDOC CRM E92 Spacetime Volume Relations
<p>This diagram illustrates the relations of the E92 Spacetime Volume entity of the CIDOC CRM ontology (http://www.cidoc-crm.org/) for Cultural Heritage Documentation.</p>
SeaLiT Knowledge Graphs - Maritime History Data in RDF using a CIDOC-CRM extension (SeaLiT Ontology)
<p><strong>SeaLiT Knowledge Graphs</strong> is an RDF dataset of maritime history data that has been transcribed (and then transformed) from original archival sources in the context of the <a href="http://www.sealitproject.eu/">SeaLiT Project</a> (Seafaring Lives in Transition, Mediterranean Maritime Labour and Shipping, 1850s-1920s). The underlying data model is the <a href="https://zenodo.org/record/5964240">SeaLiT Ontology</a>, an extension of the ISO standard <strong>CIDOC-CRM</strong> (ISO 21127:2014) for the modelling and integration of maritime history information. </p> <p>The knowledge graphs integrate data of totally 16 different types of archival sources:</p> <ul> <li>Crew Lists <ul> <li>Crew and displacement list (Roll)</li> <li>Crew List (Ruoli di Equipaggio)</li> <li>General Spanish Crew List</li> </ul> </li> <li>Registers / Lists <ul> <li>Students Register</li> <li>Civil Register</li> <li>Register of Maritime Personnel</li> <li>Register of Maritime Workers (Matricole della gente di mare)</li> <li>Sailors Register (Libro de registro de marineros)</li> <li>Naval Ship Register List</li> <li>Seagoing Personnel</li> <li>Lists of ships</li> </ul> </li> <li>Censuses <ul> <li>Census La Ciotat</li> <li>First National all-Russian Census of the Russian Empire</li> </ul> </li> <li>Payrolls <ul> <li>Payrolls of private archives and libraries in Greece</li> <li>Payrolls of Russian Steam Navigation and Trading Company</li> </ul> </li> <li>Employment records <ul> <li>Shipyards of Messageries Maritimes, La Ciotat</li> </ul> </li> </ul> <p>More information about the archival sources are available through the <a href="https://sealitproject.eu/dictionary-of-source-types-list">SeaLiT website</a>. Data exploration applications over these sources are also publicly available (<a href="https://catalogues.sealitproject.eu/">SeaLiT Catalogues</a>, <a href="http://rs.sealitproject.eu/">SeaLiT ResearchSpace</a>). </p> <p>Data from these archival sources has been transcribed in tabular form and then curated by historians of SeaLiT using the <a href="https://www.ics.forth.gr/isl/fast-cat">FAST CAT</a> system. The transcripts (records), together with the curated vocabulary terms and entity instances (ships, persons, locations, organizations), are then transformed to RDF using the SeaLiT Ontology as the target (domain) model. To this end, the corresponding schema mappings between the original schemata and the ontology were defined using the <a href="https://github.com/isl/x3ml">X3ML</a> mapping definition language, that were subsequently used for delivering the RDF datasets. </p> <p>More information about the FAST CAT system and the data transcription, curation and transformation processes can be found in the following paper:</p> <blockquote> <p>P. Fafalios, K. Petrakis, G. Samaritakis, K. Doerr, A. Kritsotaki, Y. Tzitzikas, M. Doerr, "FAST CAT: Collaborative Data Entry and Curation for Semantic Interoperability in Digital Humanities", ACM Journal on Computing and Cultural Heritage, 2021. <a href="https://doi.org/10.1145/3461460">https://doi.org/10.1145/3461460</a> [<a href="http://users.ics.forth.gr/~fafalios/files/pubs/fafaliosJOCCH2021.pdf">pdf</a>, <a href="http://users.ics.forth.gr/~fafalios/files/bibs/fafaliosJOCCH2021.bib">bib</a>]</p> </blockquote> <p>The RDF dataset is provided as a set of TriG files per record per archival source. For each record, the dataset provides: i) one trig file for the record's data (<em>records.trig</em>), ii) one trig file for the record's (curated) vocabulary terms (<em>vocabularies.trig</em>), and iii) four trig files for the record's (curated) entity instances (<em>ships.trig, persons.trig, persons.trig, organizations.trig</em>).</p> <p>We also provide the RDFS files of the used ontologies (SeaLiT Ontology verson 1.0, CIDOC-CRM version 7.1.1). </p>
SSHOC - National Gallery - Raphael Research Resource CIDOC CRM Mapped Dataset
<p>In 2007 the <a href="https://cima.ng-london.org.uk/documentation">Raphael Research Resource</a> project began to examine how complex conservation, scientific and art historical research could be combined in a flexible digital form. Exploring the presentation of interrelated high resolution images and text, along with how the data could be stored in relation to an event driven ontology in the form of <a href="http://www.w3.org/TR/rdf-concepts/">RDF triples</a>. The original <a href="https://cima.ng-london.org.uk/documentation">main user interface</a> is still live, In 2021/21 as part of the <a href="https://www.sshopencloud.eu/">SSHOC Project</a> the raw data stored within the system was mapped to the <a href="https://www.cidoc-crm.org/">CIDOC CRM</a> using a custom set of Python scripts (<a href="https://doi.org/10.5281/zenodo.6461654">https://doi.org/10.5281/zenodo.6461654</a>). The SSHOC work aimed to make this data more <a href="https://www.go-fair.org/fair-principles/">FAIR</a> so in addition to mapping it to a standard ontology, to increase Interoperability, it has also been made available in the form of <a href="http://en.wikipedia.org/wiki/Linked_Data">open linkable data</a> combined with a <a href="http://en.wikipedia.org/wiki/SPARQL">SPARQL</a> end-point. This live data presentation can been found <a href="https://rdf.ng-london.org.uk/sshoc/">Here</a>.</p> <p>This deposit contains the CIDOC-CRM mapped data formatted in XML and an example model diagram representing some of the key relationships covered in the data-set.</p>
SeaLiT Ontology - An extension of CIDOC-CRM for the modelling of Maritime History information
<p>The <strong>SeaLiT Ontology</strong> is a formal ontology intended to facilitate the integration, mediation and interchange of heterogeneous information related to <strong>maritime history</strong>. It aims at providing the semantic definitions needed to transform disparate, localised information sources of maritime history into a coherent global resource. It also serves as a common language for domain experts and IT developers to formulate requirements and to agree on system functionalities with respect to the correct handling of historical information.</p> <p>The ontology uses and extends the <strong><a href="https://www.cidoc-crm.org/">CIDOC Conceptual Reference Model</a></strong> (ISO 21127:2014), in particular version 7.2.1, as a general ontology of human activity, things and events happening in space and time.</p> <p>The ontology has been developed following a bottom-up process from primary data collected in the context of the <a href="http://www.sealitproject.eu/"><strong>SeaLiT Project</strong></a> (<em>Seafaring Lives in Transition, Mediterranean Maritime Labour and Shipping, 1850s-1920s</em>). SeaLiT is an international research project, funded by the ERC Starting Grant 2016, which explores the transition from sail to steam navigation and its effects on seafaring populations in the Mediterranean and the Black Sea between the 1850s and the 1920s.</p> <p>More information about the construction of the <strong>SeaLiT Ontology</strong>, the considered data sources, as well as their transformation to a knowledge graph using the SeaLiT Ontology, can be found in the following papers:</p> <blockquote> <p>P. Fafalios, A. Kritsotaki, and M. Doerr, "<em>The SeaLiT Ontology – An Extension of CIDOC-CRM for the Modeling and Integration of Maritime History Information"</em>. ACM Journal on Computing and Cultural Heritage, 2023. <a href="https://doi.org/10.1145/3586080">https://doi.org/10.1145/3586080</a> [<a href="https://arxiv.org/pdf/2301.04493.pdf">pdf</a>, <a href="https://users.ics.forth.gr/~fafalios/files/bibs/fafaliosSeaLiTOntology2023.bib">bib</a>]</p> </blockquote> <blockquote> <p>P. Fafalios, K. Petrakis, G. Samaritakis, K. Doerr, A. Kritsotaki, Y. Tzitzikas, and M. Doerr, "FAST CAT: Collaborative Data Entry and Curation for Semantic Interoperability in Digital Humanities", ACM Journal on Computing and Cultural Heritage, 2021. <a href="https://doi.org/10.1145/3461460">https://doi.org/10.1145/3461460</a> [<a href="http://users.ics.forth.gr/~fafalios/files/pubs/fafaliosJOCCH2021.pdf">pdf</a>, <a href="http://users.ics.forth.gr/~fafalios/files/bibs/fafaliosJOCCH2021.bib">bib</a>]</p> </blockquote> <p>The (resolvable) <strong>namespace </strong>of the ontology is: <a href="http://www.sealitproject.eu/ontology/">http://www.sealitproject.eu/ontology/</a></p> <p>An <strong>OWL implementation</strong> of the ontology is available at: <a href="https://sealitproject.eu/ontology/SeaLiT_Ontology_v1.2.owl">https://sealitproject.eu/ontology/SeaLiT_Ontology_v1.2.owl</a></p> <p><strong>Knowledge graphs </strong>that make use of the SeaLiT Ontology are available at: <a href="https://zenodo.org/record/6460841">https://zenodo.org/record/6460841</a>. These RDF datasets integrate information of 16 different types of archival sources related to maritime history, including crew lists, payrolls, registers of different types, censuses, and employment records.</p>
Dataset of KO journal paper: Semantic analysis of archival concepts in CIDOC-CRM and in RiC-CM and RiC-O
<p>Semantic analysis of archival concepts (class, relations, atributtes, and relation attributes) presents in Records in Context family (conceptual model and ontology) and its possible equivalents in CIDOC-CRM.</p>
Berliner Kunstkammer - Daten, Datenmodell und CIDOC CRM basierte Anwendungsontologie
<p><strong>*Für die Integration in den NFDI4Objects Knowledge Graphen*</strong></p> <p><strong>Übersicht der Dateien</strong></p> <p>Dieses Datenpaket enthält die Daten und Dateien, die im Kontext des DFG-Projekts "Das Fenster zur Natur und Kunst. eine historisch-kritische Aufarbeitung der Brandenburgisch-Preußischen Kunstkammer" zwischen 2018 und 2022 in der <a href="https://wiss-ki.eu/">WissKI</a>-basierten virtuellen Forschungsumgebung erschlossen bzw. zur Erschließung genutzt wurden.</p> <ul> <li>berlinerkunstkammer_daten_20240920.nq und berlinerkunstkammer_daten_20240920.nt stellen jeweils Exporte des Repositoriums (GraphDB) dar.</li> <li>berlinerkunstkammer_datenmodell_20221102T105320.xml ist der Export des WissKI Pathbuilders und repräsentiert das Datenmodell, über das die Daten semantisch erschlossen wurden</li> <li>kunstkammer.owl ist die CIDOC CRM basierte Anwendungsontologie (Erlangen CRM Version 170309), mit der das Datenmodell im WissKI Pathbuilder aufgebaut wurde.</li> </ul> <p><strong>Informationen zum Projekt</strong></p> <p>Die vom 16. bis ins 19. Jahrhundert im Berliner Schloss beheimatete Kunstkammer bildete mit ihren vielfältigen Beständen ein Fundament für zahlreiche spätere Museen. In dieser Sammlung waren Objekte der Natur, Kunst und Wissenschaft vereint. Heute sind viele der vormals der Kunstkammer zugehörenden Objekte über zahlreiche Museen Berlins verteilt. Das DFG-Projekt, eine Kooperation zwischen den Staatlichen Museen zu Berlin, der Humboldt-Universität zu Berlin und dem Museum für Naturkunde Berlin, erforschte die Sammlung und die Wege/die Geschichte ihrer Objekte bis in die heutigen Museen.</p> <p>Die Geschichte der Berliner Kunstkammer wird dabei anhand von 'Biografien' ihrer Objekte erzählt: Wie gelangten die Objekte in die Sammlung? In welchen Funktions- und Verwendungskontexten standen sie vor ihrem Eingang in die Berliner Kunstkammer? Auf welche Weise wurden die Objekte in der Sammlung in immer wieder neue taxonomische, räumliche, inszenatorische sowie nutzungsbezogene Zusammenhänge gestellt? Wie gestaltete sich ihr Eingang in die ab dem 19. Jahrhundert entstehenden Museen? Und welche Bedeutungszuweisungen gingen mit diesen Prozessen einher?</p> <p>Ausgehend von diesen Fragen wird die Kunstkammer in ihren verschiedenen historischen Transformationsstadien betrachtet. Hierzu zählen die Neuordnungsprozesse im 17. und 18. Jahrhundert, aber auch die veränderte Form der Sammlung nach Abgabe verschiedener Objektbereiche im 19. Jahrhundert. In den Blick genommen wird die enorme Bereicherung der Kunstkammer durch umfassende Privatsammlungen in dieser späten Phase ihrer Entwicklung ebenso wie die Rolle ihrer Objekte bei der Neugründung verschiedener Museen. Der objektbiografische Ansatz bildet die Grundlage, um von der Kunstkammer aus bisher kaum beachtete Facetten der Berliner Sammlungsgeschichte sichtbar zu machen und über den preußischen und Berliner Kontext hinaus neue Perspektiven für das Forschungsfeld der allgemeinen Sammlungsgeschichte zu liefern.</p> <p>Die Ergebnisse wurden in Form einer <a href="https://books.ub.uni-heidelberg.de/arthistoricum/catalog/book/1461/version/2222">Buchpublikation</a> und auf dieser <a href="https://berlinerkunstkammer.de/">virtuellen Forschungsumgebung</a> veröffentlicht.</p> <p>Die virtuelle Forschungsumgebung zur Berliner Kunstkammer übersetzt die Berliner Kunstkammer in ein digitales Wissensnetz, in dem Objekte, Akteure, Orte und Quellen miteinander verbunden sind. Sie basiert auf einer Vielzahl an unterschiedlichen Quellen (darunter Inventare, Reiseberichte, Beschreibungen), die es in ihrer Summe erlauben, eine multiperspektivische Sicht auf die Sammlung und ihre Entwicklung zu gewinnen. </p> <p>Zahlreiche Objekte der Berliner Kunstkammer sind heute zwar noch in den Sammlungen der <a href="https://www.smb.museum/home/">Staatlichen Museen zu Berlin</a> erhalten oder sind in die Bestände des <a href="https://www.museumfuernaturkunde.berlin/de">Museums für Naturkunde Berlin</a> eingegangen, andere werden in jenen der <a href="https://www.hu-berlin.de/de">Humboldt-Universität zu Berlin</a> aufbewahrt. Ein großer Teil der Objekte aber ist heute nur noch in den historischen Quellen überliefert, weshalb der Fokus der Forschungsumgebung auf einer textbasierten Bestandsrekonstruktion liegt, um die Objekte digital und in Zeitschichten rekonstruiert recherchierbar zu machen. Neben einer <a href="https://berlinerkunstkammer.de/uebersicht-der-quellen">Übersicht</a> zu zentralen Quellen zur Brandenburgisch-Preußischen Kunstkammer und ihren Vorgänger- und Nachfolgeinstitutionen werden ausgewählte Quellen mit Transkriptionen (Arbeitsversionen) und Digitalisaten sowie mit tiefenerschlossenen Objektinformationen Forschenden zur Verfügung gestellt.</p> <p><strong>Disclaimer Sensible Inhalte</strong></p> <p>Diese Forschungsumgebung enthält Ressourcen mit heute unangemessenen Begriffen und Beschreibungen aus historischen Quellen, die rassistische, diskriminierende und abwertende Sprache enthalten. Das Projekt bemüht sich darum, auch diese historische Dimension der Sammlungen sichtbar zu machen und sich damit kritisch auseinanderzusetzen. Die Mitarbeiter*innen des Projekts distanzieren sich von diesem Sprachgebrauch und den in den Materialien und Quellen gespiegelten Ansichten.</p> <p><strong>Technische Architektur der Forschungsumgebung</strong></p> <p>Die Besonderheit der virtuellen Berliner Kunstkammer liegt in der Art der Beschreibung, Vernetzung und Abbildung ihrer Inhalte. Sie basiert auf Technologien des <a href="https://www.w3.org/standards/semanticweb"><em>Semantic Web</em></a>, bei denen Funktionalitäten des World Wide Web um eine zusätzliche Bedeutungsebene erweitert werden, um Wissen zu strukturieren und besser verwert- und austauschbar zu machen. Das Grundgerüst der virtuellen Berliner Kunstkammer bildet die Forschungs- und Dokumentationsumgebung <a href="http://wiss-ki.eu/">WissKI</a>, die auf die webbasierte, semantische Erschließung, Erforschung und Publikation kulturellen Erbes ausgerichtet ist. Die Topologie der virtuellen Berliner Kunstkammer ist durch das ihr zugrundeliegende Datenmodell definiert, das den Forschungsgegenstand und die Fragestellungen an diesen in einen Wissensgraphen überträgt. Zentral bei der Erforschung dieser historischen Sammlung sind die Wege und der Bedeutungswandel ihrer Objekte und die damit verbundenen Auswirkungen auf die Konstellation und Dynamik der Bestände unter Berücksichtigung der Historizität der Quellen und der damit verbundenen Mehrdeutigkeit, aber auch Lückenhaftigkeit von Information. Die Abdeckung all dieser Aspekte geht zudem einher mit dem Anspruch an eine Erweiterbarkeit der Inhalte aus den trotz Kriegsverlusten umfangreich überlieferten Archivalien aus mehreren Jahrhunderten.</p> <p>Ein solcher Wissensgraph repräsentiert Informationen eines Gegenstandsbereichs, die mit Ontologien – in der Informatik eine Sammlung an Begriffen und der zwischen ihnen bestehenden Beziehungen – formal beschrieben werden. Für das Projekt wurde eine auf dem <a href="http://www.cidoc-crm.org/">CIDOC Conceptual Reference Model (CRM)</a> bzw. seiner <a href="http://erlangen-crm.org/">OWL-Implementierung</a> basierende Anwendungsontologie geschaffen. Diese Ontologie, die auch in diesem Projekt zum Einsatz kam, besteht aus ca. 90 Klassen (z. B.<em> Physischer Gegenstand, Akteur</em>, <em>Ort</em>, <em>Zeitspanne</em>) und 150 Beziehungen (z. B. <em>hat name</em>, <em>fand statt in</em>, <em>wurde erschaffen durch</em>), deren jeweilige Bedeutung genau definiert ist. Sie wird von Geisteswissenschafler*innen und Informatiker*innen stetig weiterentwickelt und dient als <em>Lingua Franca</em> für transdisziplinäre Forschung. Durch die Verkettung von Klasse – Beziehung – Klasse werden Sätze gebildet, die von Mensch und Maschine verwertet werden können. All diese Sätze eines betrachteten Gegenstandsbereichs verknüpft und im Kontext spezifischer Fragestellungen miteinander in Verbindung gebracht, entsteht ein auf den Anwendungsbereich zugeschnittenes, graphbasiertes Datenmodell bzw. ein Wissensnetz. Als ereigniszentrierte Ontologie bietet das CIDOC CRM zudem die Möglichkeit, Zustände und ihre Veränderung über Ereignisse (z. B. <em>Geburt</em>, <em>Herstellung</em>, <em>Zuweisung</em>) abzubilden, an die wiederum Akteure, Zeit oder Ort über entsprechende Relationen gekoppelt und Zustandsveränderungen somit kontextualisiert werden können. Somit bietet diese Ontologie ideale Voraussetzungen zur quellenbasierten Rekonstruktion historischer Sammlungsbestände.</p> <p> </p> <p>Quelle: <a href="https://berlinerkunstkammer.de/vfu">https://berlinerkunstkammer.de/vfu</a> und <a href="https://berlinerkunstkammer.de/projekt">https://berlinerkunstkammer.de/projekt</a></p> <p>Weiterführende Literatur: <a href="https://books.ub.uni-heidelberg.de/arthistoricum/catalog/book/1461/chapter/21156">Sarah Wagner: Vom Schloss ins Internet - Die virtuelle Forschungsumgebung zur Berliner Kunstkammer. In: Becker, Marcus et al. (Hrsg.): Die Berliner Kunstkammer: Sammlungsgeschichte in Objektbiografien vom 16. bis 21. Jahrhundert, Heidelberg: arthistoricum.net-ART-Books, 2024</a>.</p> <p> </p>
APIS CIDOC CRM serialization
<p>This is a trig serialisation of the enriched version of the Austrian Biographical Dictionary. The data was enriched during the Austrian Prosopographical Information System (APIS) project. It contains data on ~19.000 persons who had an impact on Austrian soil and died between 1815 and 1955.</p> <p>This version of the APIS data uses CIDOC CRM and is less complete than the also available JSON version of the data (using a custom data model).</p> <p>Please refer to <a href="https://www.oeaw.ac.at/acdh/team/current-team/matthias-schloegl/">Matthias Schlögl</a> for any inquiries.</p>
SSHOC - National Gallery - Grounds Database CIDOC CRM Mapped Dataset
<p>In 2018 the <a href="https://doi.org/10.5281/zenodo.5838339">IPERION-CH Grounds Database</a> was presented to examine how the data produced through the scientific examination of historic painting preparation or grounds samples, from multiple institutions could be combined in a flexible digital form. Exploring the presentation of interrelated high resolution images, text, complex metadata and procedural documentation. The original <a href="https://research.ng-london.org.uk/iperion/">main user interface</a> is live, though password protected at this time. Work within the <a href="https://www.sshopencloud.eu/">SSHOC project</a> aimed to reformat the data to create a more <a href="https://www.go-fair.org/fair-principles/">FAIR</a> data-set, so in addition to mapping it to a standard ontology, to increase Interoperability, it has also been made available in the form of <a href="http://en.wikipedia.org/wiki/Linked_Data">open linkable data</a> combined with a <a href="http://en.wikipedia.org/wiki/SPARQL">SPARQL</a> end-point. A draft version of this live data presentation can been found <a href="https://rdf.ng-london.org.uk/sshoc/">Here</a>.</p> <p>This is a draft data-set and further work is planned to debug and improve its semantic structure.This deposit contains the CIDOC-CRM mapped data formatted in XML and an example model diagram representing some of the key relationships covered in the data-set.</p>
APIS Person modeled in CIDOC CRM
<p>Structured data of a APIS/ÖBL biography modeled in CIDOC CRM, serialized in turtle.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.