Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

266

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

266 results for “xml”

Learn how ShareScore rates datasets ↗
zenodo44/100

IN02067 Inscription of Bhimarjuna and Visnugupta at Yengahiti. Sanskrit XML file, draft epidoc edition

<p>IN02067 Inscription of Bhimarjuna and Visnugupta at Yengahiti. Sanskrit XML file (without metadata). Draft epidoc edition to be incorporated into &#39;Siddham&#39; archive</p>

opencc-by-4.0Mar 2018View details →
zenodo44/100

IN02080 Yengu Bahaltole Inscription. Sanskrit XML file, draft epidoc edition

<p>IN02080 Yengu Bahaltole Inscription. Sanskrit XML file (without metadata). Draft epidoc edition to be incorporated into &#39;Siddham&#39; archive</p>

opencc-by-4.0Mar 2018View details →
zenodo44/100

IN02069 Tebahal Stone Inscription. Sanskrit XML file, draft epidoc edition

<p>IN02069 Tebahal Stone Inscription. Sanskrit XML file (without metadata). Draft epidoc edition to be incorporated into &#39;Siddham&#39; archive</p>

opencc-by-4.0Mar 2018View details →
zenodo44/100

IN02081 Sanku Fragment Inscription (revise title). Sanskrit XML file, draft epidoc edition

<p>IN02081 Sanku Fragment Inscription (revise title). Sanskrit XML file (without metadata). Draft epidoc edition to be incorporated into &#39;Siddham&#39; archive</p>

opencc-by-4.0Mar 2018View details →
zenodo44/100

Inundation maps of Donana for 23 dates within the period 2015/12/19 to 2017/08/20 and their accompanying INSPIRE metadata XML files

<p>Satellite-derived inundation maps offer an efficient solution for monitoring the spatial and temporal variability of the hydrological cycle of wetlands. This task is important for taking mitigation actions against factors (e.g. climate change and human pressures) threatening wetlands&#39; functions and services.</p> <p>Inundation maps&nbsp;within the period 2015/12/19 to 2017/08/20 were generated for Donana based on the methodology presented in &quot;Kordelas, G.A.; Manakos, I.; Aragon&eacute;s, D.; D&iacute;az-Delgado, R.; Bustamante, J. Fast and Automatic Data-Driven Thresholding for Inundation Mapping with Sentinel-2 Data. <em>Remote Sens.</em> <strong>2018</strong>, <em>10</em>, 910.&quot;.</p> <p>Each inundation map is named as &quot; &#39;Date&#39;_inundation_map_Donana_S2.tif &quot;, and contains the following classes: Inundated Class, Non-inundated Class. In this map, Inundated and Non-inundated Classes are denoted with 0 and 1, respectively. &#39;Date&#39; is in the form YYYY_MM_DD.</p>

opencc-by-4.0Sep 2019View details →
zenodo44/100

Inundation maps of Danube Delta for 10 dates within the period 2016/10/05 to 2017/08/01 and their accompanying INSPIRE metadata XML files

<p>Satellite-derived inundation maps offer an efficient solution for monitoring the spatial and temporal variability of the hydrological cycle of wetlands. This task is important for taking mitigation actions against factors (e.g. climate change and human pressures) threatening wetlands&#39; functions and services.</p> <p>Inundation maps&nbsp;within the period 2016/10/05 to 2017/08/01 were generated for Danube Delta based on the methodology presented in &quot;Kordelas, G.A.; Manakos, I.; Aragon&eacute;s, D.; D&iacute;az-Delgado, R.; Bustamante, J. Fast and Automatic Data-Driven Thresholding for Inundation Mapping with Sentinel-2 Data. <em>Remote Sens.</em> <strong>2018</strong>, <em>10</em>, 910.&quot;.</p> <p>Each inundation map is named as &quot; &#39;Date&#39;_inundation_map_Danube_Delta_S2.tif &quot;, and contains the following classes: Inundated Class, Non-inundated Class. In this map, Inundated and Non-inundated Classes are denoted with 0 and 1, respectively.&nbsp;The regions, which are manually denoted as affected by clouds, are denoted with 2. &#39;Date&#39; is in the form YYYY_MM_DD.</p>

opencc-by-4.0Sep 2019View details →
zenodo44/100

Inundation maps of Camargue for 47 dates within the period 2016/02/09 to 2018/06/19 and their accompanying INSPIRE metadata XML files

<p>Satellite-derived inundation maps offer an efficient solution for monitoring the spatial and temporal variability of the hydrological cycle of wetlands. This task is important for taking mitigation actions against factors (e.g. climate change and human pressures) threatening wetlands&#39; functions and services.</p> <p>Inundation maps&nbsp;within the period 2016/02/09 to 2018/06/19 were generated for Camargue based on the methodology presented in &quot;Kordelas, G.A.; Manakos, I.; Aragon&eacute;s, D.; D&iacute;az-Delgado, R.; Bustamante, J. Fast and Automatic Data-Driven Thresholding for Inundation Mapping with Sentinel-2 Data. <em>Remote Sens.</em> <strong>2018</strong>, <em>10</em>, 910.&quot;.</p> <p>Each inundation map is named as &quot; &#39;Date&#39;_inundation_map_Camargue_S2.tif &quot;, and contains the following classes: Inundated Class, Non-inundated Class. In this map, Inundated and Non-inundated Classes are denoted with 0 and 1, respectively. &#39;Date&#39; is in the form YYYY_MM_DD.</p>

opencc-by-4.0Sep 2019View details →
zenodo44/100

AquaMaps: AquaMaps XML resource

from <p></p>http://www.aquamaps.org/. AquaMaps are computer-generated predictions of natural occurrence of marine species, based on the environmental tolerance of a given species with respect to depth, salinity, temperature, primary productivity, and its association with sea ice or coastal areas. These __environmental envelopes__ are matched against an authority file which contains respective information for the Oceans of the World. Independent knowledge such as distribution by FAO areas or bounding boxes are used to avoid mapping species in areas that contain suitable habitat, but are not occupied by the species. Maps show the color-coded likelihood of a species to occur in a half-degree cell, with about 50 km side length near the equator. Experts are able to review, modify and approve maps.<p></p>from EOL v2 database

opencc-by-4.0Aug 2024View details →
zenodo44/100

AnAge: AnAge text (XML resource)

AnAge is a database of longevity and ageing in animals. It features quantitative life history data for over 4,000 species, including extensive longevity records, body masses at different developmental stages, reproductive data, and physiological traits related to metabolism. In addition to quantitative data, AnAge also features comments and observations related to ageing or relevant to the life history of individual taxa. AnAge features a manually-curated collection of animal longevity records. Moreover, AnAge has extensive life-history traits such as adult body weight, gestation or incubation time, age at sexual maturity and other reproductive data. Lastly, observations on physiological or pathological changes with age in animals are (where available) featured. AnAge focuses primarily on chordates. At the time of writing, AnAge features 4,122 entries, including 1,331 mammals, 1,098 birds, 539 reptiles, 169 amphibians, 962 fishes and 28 non-chordates. Our focus is on accuracy and quality, however, not quantity, and only species for which we have confidence in the data are featured. Numerous experts have contributed information to AnAge and helped us meet quality standards. Professor Steven Austad, a world-renowned expert in mammalian ageing at the Barshop Institute in San Antonio, is AnAge__s expert mammalogist and curator.<p></p>

opencc-by-4.0Aug 2024View details →
zenodo44/100

AskNature: AskNature XML

Open the record for dataset details and reuse information.

opencc-by-4.0Aug 2024View details →
zenodo44/100

TEI-XML Zürcher Regierungsratsbeschlüsse 1803-1887

<p><strong>Projekt TKR</strong></p> <p>Zwischen 2003&ndash;2016 wurden die Protokollb&auml;nde des Regierungsrats und des Kantonsrats als Worddokumente seriell transkribiert und anschliessend in PDF und sp&auml;ter als OGD-Datens&auml;tze in TEI-XML konvertiert.<br>Die Dateien werden unter Ber&uuml;cksichtigung der gesetzlichen Schutzfristen laufend (max. 80 Jahre) publiziert.</p> <p>Siehe auch: <a href="https://archives-quickaccess.ch/search/stazh/rrb">https://archives-quickaccess.ch/search/stazh/rrb</a></p> <p><strong>Inhalt</strong></p> <p>Dieses Datenset beinhaltet die <strong>handschriftlich verfassten Regierungsratsbeschl&uuml;sse des Kantons Z&uuml;rich von 1803 bis 1887</strong>. Ab dem 1. Juli 1887 wurden die Regierungsratsbeschl&uuml;sse gedruckt.&nbsp;</p> <p>Die Beschl&uuml;sse des Regierungsrates geh&ouml;ren zu den zentralen Aktenserien des Kantons Z&uuml;rich.&nbsp;<br>In ihnen spiegelt sich ein &auml;usserst breites Spektrum an Themen, da der Regierungsrat nicht nur f&uuml;r grosse politische Entscheide, sondern oft auch f&uuml;r allt&auml;gliche Belange zust&auml;ndig war.&nbsp;<br>Neben Themen wie Auswanderung oder Aufnahme von politischen Fl&uuml;chtlingen, Bau von Eisenbahnen und Strassen oder Regulierung der stark ansteigenden Industrie besch&auml;ftigen den Regierungsrat stets auch Tagesgesch&auml;fte wie Konzessionsgesuche f&uuml;r Tavernen oder Wasserkraftanlagen, Steuerrekurse, Einb&uuml;rgerungen oder die Aufnahme von Kantonsfremden in kantonseigene Spit&auml;ler.<br>Die Metadaten eines Beschlusses (TEI/teiHeader) bestehen unter anderem aus dem K&uuml;rzel des/der Transkriptors/in, dem Transkriptionsdatum, der Signatur, dem Publikationsdatum, dem Titel des Beschlusses, dem Link zur Verzeichnung im Archivkatalog und dem Beschlussdatum.<br>Klassen und B&auml;nde bilden die hierarchische Ordnerstruktur.<br>Eine Klasse stellt jeweils den Zeitraum zwischen zwei politischen Umbr&uuml;chen dar:&nbsp;<br>Das Protokoll setzt 1803 mit der Bildung des modernen Kantons Z&uuml;rich ein, die erste Klasse schliesst mit der Annahme der neuen restaurativen Verfassung im Juni 1814.&nbsp;<br>Die dritte Klasse setzt mit der liberalen Verfassung von 1831 ein; der Bruch ist auch daran erkennbar, dass der &laquo;Kleine Rat&raquo; von nun an &laquo;Regierungsrat&raquo; genannt wird.&nbsp;<br>Mit dem konservativen &laquo;Z&uuml;riputsch&raquo; im September 1839 beginnt die vierte Klasse, welche 1849 mit einer gross angelegten Verwaltungsreform und der Integration des Kantons Z&uuml;rich in den neuen schweizerischen Bundesstaat schliesst.&nbsp;<br>Mit der demokratischen Verfassung von 1869 setzt die sechste und letzte Klasse ein, welche nicht durch ein politisches Ereignis, sondern durch die Umstellung auf das gedruckte Protokoll endet.&nbsp;<br>Dies entspricht auch der Struktur im Archivverzeichnis.</p> <p><strong>Technische Erschliessung</strong></p> <p>Die Konvertierung von Worddokumenten in TEI-konforme XML-Dateien geschah mittels eines Python-Scripts, welches mittels Mustererkennung die einzelnen Datenelemente voneinander abgrenzte.&nbsp;</p> <p>Validierung:</p> <p>Alle Dateien sind gem&auml;ss TEI-Schema valide. Nicht valide xml-Dateien wurden manuell verbessert.</p> <p>XSL-Stylesheet:</p> <p>Das Stylesheet befindet sich im Ordner Ressourcen ("\Ressourcen\Stylesheet.xsl") und ist mit relativem Pfad eingebunden in den XML-Dateien (Zeile 2: &lt;?xml-stylesheet type="text/xsl" href="../../Ressourcen/Stylesheet.xsl"?&gt;).&nbsp;<br>Das Stylesheet erm&ouml;glicht eine Browseransicht, welche sich der originalen Transkription in Word ann&auml;hert.</p> <p>Ab Version 3.0 wurden im ganzen Datensatz Eigennamen mittels maschinellem Lernen ausgezeichnet (vgl. <a href="https://github.com/machinelearningZH/named-entity-recognition_staatsarchiv">Github-Repository</a>&nbsp;zum Projekt).&nbsp;<em>Bitte beachten:&nbsp;</em>Die Auszeichnung wurde nicht manuell nachkontrolliert und kann Fehler enthalten!</p> <p><strong>Urheberrecht</strong></p> <p>Die Daten stehen unter einer Creative Commons CC-BY-SA 4.0 Lizenz.</p> <p>Herausgeber: Staatsarchiv des Kantons Z&uuml;rich</p> <p>Technische Erschliessung: Rebekka Pl&uuml;ss, rebekka.pluess@zh.ch&nbsp;</p> <p>Projektleiter TKR: Luzi Schutz</p>

opencc-by-sa-4.0Jun 2017View details →
zenodo40/100

morethanbooks/XML-TEI-Bible: XML-TEI Bible: Entities and communication (66 Books)

<p>This release contains the biblical text in XML-TEI (66 books). The encoded text is in Spanish, but the codification (elements, attributes, values, ids) is in English. It makes explicit following information:</p> <ul> <li>Books, chapters, pericopes and verses.</li> <li>References to peoples, places, times, groups and books, using ids.</li> <li>Direct speech, including who is communicating, to whom and how (written, oral, prayer...).</li> </ul>

openother-openMay 2020View details →
zenodo40/100

Ancient Greek Literature for Advanced Data Processing: A Text Fabric Representation of Open Access Texts in TEI XML

<p>This data set contains a full conversion of Greek texts available in the Perseus Digital Library and the Open Greek and Latin Project to the Text Fabric data format. The main advantage of the Text Fabric datatype over the original TEI XML format is that it utilizes a strict separation of text and annotation in a flat data structure. At the same time, it permits multiple distinct formats of the same text as well as an unlimited depth of (embedded) annotations. Because of its flat data structure, it facilitates easy and clean procedures to analyze, transform, and enrich the available data. Many of these processes are very difficult to conduct while departing from the hierarchically organized XML tree representation.</p>

opencc-by-4.0Dec 2020View details →
zenodo40/100

XML_corpus

<p>All texts are from TextGrid licenced under CC-BY 3.0 (https://creativecommons.org/licenses/by/3.0/de/) and put together as a corpus by Dr. Katrin Dennerlein (http://www.germanistik.uni-wuerzburg.de/lehrstuehle/computerphilologie/mitarbeiter/dennerlein/).</p> <p>The corpus is mentioned in The Schiller-Kleist Uncertainty Principle.</p>

opencc-by-4.0Jul 2014View details →
zenodo40/100

A word2vec model file built from the French Wikipedia XML Dump using gensim.

<p>A word2vec model file built from the French Wikipedia XML dump using gensim. The data published here includes three model files (you need all three of them in the same folder) as well as the Python script used to build the model (for documentation). The Wikipedia dump was downloaded on October 7, 2016 from https://dumps.wikimedia.org/. Before building the model, plain text was extracted from the dump. The size of that dataset is about 500 million words or 3.6 GB of plain text. The principal parameters for building the model were the following: no lemmatization was performed, tokenization was done using the "\W" regular expression (any non-word character splits tokens), and the model was built with 500 dimensions.</p>

opencc-by-4.0Oct 2016View details →
zenodo40/100

TEI-XML-Datenset der Tagebücher, Briefe, Dokumente, Forschungsbeiträge, Chronologieeinträge und Register der edition humboldt digital

<p>Das Datenset enth&auml;lt alle edierten Texte (Tageb&uuml;cher, Briefe und weitere Dokumente) sowie Paratexte (Forschungsbeitr&auml;ge, Eintr&auml;ge der Chronologie zu Alexander von Humboldts Leben, Register und Glossar) der Version 11 der <a href="https://edition-humboldt.de">edition humboldt digital</a>, die am 4. Juni 2025 erschienen ist. Das Datenset enth&auml;lt gegen&uuml;ber der HTML-Version technische Fehlerkorrekturen, daher wird es als Version 11.0.1 ver&ouml;ffentlicht.</p> <p>Die Editionsrichtlinien stehen auf <a href="https://edition-humboldt.de/richtlinien/index.html">edition-humboldt.de</a> zur Vef&uuml;gung. Das Datenmodell ist in drei verschiedene ODDs aufgeteilt (f&uuml;r edierte Texte, Registereintr&auml;ge und Forschungsbeitr&auml;ge). Dem Datenset liegen die drei RNG-Schemata bei, die ODD-Ursprungsdateien sind im GitHub-Repository <a href="https://github.com/telota/ediarum.AVHR.data-model/">ediarum.AVHR.data-model</a> zu finden. Beachten Sie bitte, dass es f&uuml;r das Pflanzenregister derzeit noch kein Schema gibt, da dieses aus dem Tagging automatisch erstellt wird.</p> <p>Weitere Hinweise zur digitalen Methodik finden sich in <a href="https://edition-humboldt.de/H0016212">Dumont 2024</a> und zum Editionsvorhaben im Allgemeinen in <a href="https://doi.org/10.25365/wdr-01-03-02">Kraft/Dumont 2020</a>.</p> <p>Dieses Datenset ist auch auf <a href="https://github.com/telota/edition-humboldt-digital">GitHub</a> zug&auml;nglich.</p>

opencc-by-sa-4.0Jul 2023View details →
zenodo40/100

XML-Schema (Gemein-Nachrichten)

<p>Mit den&nbsp;<em><strong>Gemein-Nachrichten</strong></em>&nbsp;stellt das Unit&auml;tsarchiv Herrnhut der weltweiten Evangelischen Br&uuml;der-Unit&auml;t - Herrnhuter Br&uuml;dergemeine (Unitas Fratrum / Moravian Church) das &auml;lteste und umfangreichste Mitteilungsblatt der Br&uuml;dergemeine digital zur Verf&uuml;gung. Es enth&auml;lt Berichte aus Gemeinden sowie dem Missions- und Diasporawerk der Br&uuml;dergemeine sowie Reden und Lebensl&auml;ufe. Die&nbsp;<em>Gemein-Nachrichten</em>&nbsp;wurden ab 1765 in Fortsetzung des&nbsp;<em>J&uuml;ngerhaus-Diariums</em>&nbsp;(1747-1764) ausschlie&szlig;lich handschriftlich vervielf&auml;ltigt. In Druck gingen 1817 und 1818 die&nbsp;<em>Beytr&auml;ge aus der Br&uuml;der-Gemeine</em>&nbsp;und zwischen 1819 und 1894 die&nbsp;<em>Nachrichten aus der Br&uuml;der-Gemeine</em>. Das Nachrichtenblatt fand in den&nbsp;<em>Mitteilungen aus der Br&uuml;der-Gemeine zur F&ouml;rderung christlicher Gemeinschaft</em>&nbsp;ab 1895 bis 1941 seine Fortsetzung.</p> <p>Mit dem hier vorliegenden <strong>XML-Schema</strong> werden erschlossene Transkripte mit standardisierten Metadaten angereichert.</p>

opencc-by-4.0Jan 2024View details →
zenodo40/100

IN01055 Halsi Grant of Ravivarman (5 plates). Sanskrit XML file

<p><a href="https://siddham.network/inscription/in01055/">IN01055</a> Halsi Grant of Ravivarman (5 plates). Sanskrit XML file (without metadata).</p>

opencc-by-4.0Jul 1996View details →
zenodo40/100

IN01048 Banavasi Inscription of Mrgesavarman. Sanskrit XML file

<p>IN01048 Banavāsi Inscription of Mṛgeśavarman. Sanskrit XML file (without metadata).</p>

opencc-by-4.0Jul 1996View details →
zenodo40/100

IN01062 Sivalli Grant of Krsnavarman II, Year 7. Sanskrit XML file

<p>IN01062 Śivaḷḷi Grant of Kṛṣṇavarman II, Year 7. Sanskrit XML file (without metadata).</p>

opencc-by-4.0Jul 1996View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record