Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

63

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

63 results for “charters”

Learn how ShareScore rates datasets ↗
zenodo44/100

Late Latin Charter Treebank 1 (LLCT1), version 1.2

<p>Version 1.2 of the Late Latin Charter Treebank 1 (LLCT1). Contains a number of minor corrections, replaces the version 1.0 published at Zenodo in 2018. Early Medieval Latin documentary texts from Italy between AD 714-869 with morphological and syntactic annotation. Latin Dependency Treebank (LDT) compatible linguistic annotation, Prague style treebank format (PML). For a detailed description of the Late Latin Charter Treebanks, see the pre-print of the paper &#39;Late Latin Charter Treebank: contents and annotation&#39;, to be published in Corpora, 16:2 (2021), at the <a href="https://researchportal.helsinki.fi/fi/publications/late-latin-charter-treebank-contents-and-annotation">institutional repository of the University of Helsinki</a>. See also Korkiakangas, T. and Lassila, M. (2013), <a href="https://www.academia.edu/5491302/Korkiakangas_Timo_and_Lassila_Matti_Abbreviations_fragmentary_words_formulaic_language_treebanking_mediaeval_charter_material_"><em>Abbreviations, fragmentary words, formulaic language: treebanking medieval charter material</em></a>, in Mambrini, F., Passarotti, M. and Sporleder, C., <em>Proceedings of the third workshop on annotation of corpora for research in the humanities</em>, pp. 61&ndash;72, and Korkiakangas, T. and Passarotti, M. (2011), <a href="https://pdfs.semanticscholar.org/6825/a8ad70fe6e2a77540d9ff2774b8f34804fd0.pdf?_ga=2.234021022.1247020789.1580545874-419947753.1580545874"><em>Challenges in Annotating Medieval Latin Charters</em></a>, in &laquo;Journal of Language Technology and Computational Linguistics&raquo;, 26, pp. 103&ndash;114.</p>

opencc-by-4.0Jan 2020View details →
zenodo44/100

Jagjivanpur, West Bengal. Detail of the copper-plate charter of Mahendrapāla

<p>Jagjivanpur, West Bengal. Detail of the copper-plate charter of Mahendrapāla, now in the collection of the Malda Museum, Malda, West Bengal, as documented in 2007.</p>

opencc-by-4.0Jun 2017View details →
zenodo44/100

Jagjivanpur, West Bengal. Detail of the copper-plate charter of Mahendrapāla

<p>Jagjivanpur, West Bengal. Detail of the copper-plate charter of Mahendrapāla, now in the collection of the Malda Museum, Malda, West Bengal, as documented in 2007.</p>

opencc-by-4.0Jun 2017View details →
zenodo44/100

Charters and Records of Königsfelden Abbey and Bailiwick (1308-1662)

<p>The data has been published online as a scholarly edition: <a href="https://www.koenigsfelden.uzh.ch/">www.koenigsfelden.uzh.ch</a>.</p> <p>The charters and records have been digitized in cooperation with the <a href="https://www.ag.ch/de/bks/kultur/archiv_bibliothek/staatsarchiv/staatsarchiv.jsp">State Archives of the Aargau</a> (StAAG). All images are available in public domain. The cartularies are available via <a href="http://e-codices.ch/en/search/?aSelectedFacets=%7B%22collection_facet%22%3A%5B%22Aarau%2C+Staatsarchiv+Aargau%22%5D%7D&amp;sQueryString=cartulary&amp;sSearchField=fullText&amp;sSortField=score">e-codices</a>.</p> <p>The data set has been manually prepared as a scholarly edition with information about layout and text. The data is available as <a href="https://ocr-d.de/de/gt-guidelines/trans/trPage.html">PageXML</a> and as TEI XML, prepared according to the <a href="https://tei-c.org/">TEI</a> (Text Encoding Initiative, specified by the Swiss Law Sources: <a href="https://www.ssrq-sds-fds.ch/wiki/">www.ssrq-sds-fds.ch/wiki/</a>). Further information about the scholarly edition can be found online: <a href="https://www.koenigsfelden.uzh.ch/exist/apps/ssrq/intro.html">www.koenigsfelden.uzh.ch/exist/apps/ssrq/intro.html</a>.</p> <p>The PageXML are structured in 28 collections. The identification of TEI to PageXML (and back) is given by the file names.</p> <p>The textual data is licensed under a <a href="https://creativecommons.org/licenses/by/4.0/">CC-BY 4.0 International</a> license.</p> <p>The data set is split into</p> <ul> <li>all images (except cartularies) as JPG</li> <li>all files in TEI XML</li> <li>PageXML of charters and records</li> <li>PageXML of the cartularies</li> </ul> <p>[German Abstract]</p> <p>Das Projekt &laquo;Urkunden und Akten des Klosters und Oberamts K&ouml;nigsfelden&raquo; will den historischen K&ouml;nigsfelder Urkunden- und Aktenbestand aus der Zeit bis 1662 online und als print-on-demand Buch edieren und das gesamte kl&ouml;sterliche Verwaltungsschriftgut digital zug&auml;nglich machen. Die Arbeiten erfolgen in Zusammenarbeit mit dem Staatsarchiv des Kantons Aargau und sind auf vier Jahre angelegt. Mit Unterst&uuml;tzung des Z&uuml;rcher Rechtsquellenprojekts der Schweizerischen Rechtsquellenstiftung streben wir eine digitale Edition an, welche die wesentlichen Vorteile aktueller technischer M&ouml;glichkeiten nutzt, die Dokumente ad&auml;quat auszeichnet, gezielt miteinander verkn&uuml;pft und f&uuml;r neue Zugriffsm&ouml;glichkeiten aufbereitet. Das R&uuml;ckgrat des Unternehmens bildet der verh&auml;ltnism&auml;ssig geschlossene Bestand der &uuml;berlieferten mittelalterlichen und fr&uuml;hneuzeitlichen Einzelblattdokumente des Klosters (STAAG U.17, 1291-1789), der - erweitert um ausw&auml;rtige St&uuml;cke aus dem einstigen Kloster - in seiner historischen Entwicklung und in seinen wandelbaren Ordnungen nachvollziehbar gemacht werden soll. Aufgrund von nachtr&auml;glich auf den Urkunden aufgebrachten Dorsualnotizen und Signaturen sollen historische Archivordnungen rekonstruiert sowie die sp&auml;teren Abschriften dieser Urkunden nachgewiesen und verlinkt werden. So kann die Edition zugleich die Umorganisation und Umdeutung dieses Bestandes erschliessen. Wir nutzen die besondere Beweglichkeit einer digitalen Edition, um der Forschung in neuartiger Weise Entwicklungen eines Bestandes und dahinterliegende Umbr&uuml;che in der Schrift- und Administrationskultur zug&auml;nglich zu machen. Zugleich soll das Projekt M&ouml;glichkeiten digitalen Edierens im fachlichen Austausch reflektieren und vorantreiben.Mit dem historisch herausragenden Dokumenten- und Kopialbuchbestand der habsburgischen Klosterstiftung und der sp&auml;teren Berner Landvogtei stellt die Edition bislang weitgehend fehlende Arbeitsgrundlagen f&uuml;r aktuelle Forschungsrichtungen bereit. Die K&ouml;nigsfelder Dokumente sind von h&ouml;chstem Interesse f&uuml;r die Kunst-, Kultur-, Sozial- und Wirtschaftsgeschichte (vom Klosteralltag &uuml;ber den Klosterhaushalt bis hin zu Hofhaltung und Armenf&uuml;rsorge), f&uuml;r die historischen Gender Studies und die Ordens- und Reformationsgeschichte. Zugleich wird unsere Edition neuen Ans&auml;tzen der Schweizer Geschichte entgegenkommen. Denn sie macht Gemengelagen zwischen unterschiedlichen Herrschaften (Habsburg, Kloster als Territorialherr, eidgen&ouml;ssische Orte) ebenso wie lokale Dimensionen von Staatsbildung, Territorialisierung und Konfessionalisierung fassbar. Das Projekt wird an der Universit&auml;t Z&uuml;rich angesiedelt, zielt auf eine enge Verkopplung von Editionsarbeit mit wissenschaftlicher Forschung und Lehre und wird Studierende mit Editionstechniken und archivalischen Gegebenheiten vertraut machen. Ausserdem wird das Projekt in Kooperation mit dem Staatsarchiv Aargau, dem Museum Aargau sowie weiteren Institutionen einen Beitrag zur Geschichtsvermittlung und Kulturg&uuml;terpflege leisten.</p> <p>Das Projekt&hellip;</p> <ul> <li>rekonstruiert und ediert einen historischen Dokumentenbestand und erschliesst weitere bislang kaum edierte Typen kl&ouml;sterlicher und herrschaftlicher Quellen</li> <li>unterst&uuml;tzt neue, &uuml;ber das &laquo;Werden und Wachsen der Eidgenossenschaft&raquo; hinausblickende Zug&auml;nge zur Geschichte der Schweiz</li> <li>leistet einen Beitrag zur Weiterentwicklung der offenen digitalen Edition</li> <li>wendet sich mit akademischer Lehre und &uuml;ber Kooperationen direkt an Studierende und ein gr&ouml;sseres Publikum</li> </ul>

opencc-by-4.0Aug 2021View details →
zenodo44/100

MPS Data set with images of medieval charters for handwriting-style based dating of manuscripts

<pre>The MPS benchmark data set for handwritten manuscript dating ____________________________________________________________ This data set is collected for the Dutch NWO project: Medieval Paleographical Scale (MPS) by Petros Samara Project website: http://application02.target.rug.nl/monk/Projects/MPS/ Copyright (c) Huygensinstituut, Den Haag, 2016 University of Groningen, 2016. All rights reserved. Organisation of the data: Each .tar.gz file contains a number of NetPBM images. The format is chosen because of its simplicity. Also, there is no doubt about lossy compression in the processing chain. The file names are of the format &#39;MPS&lt;year&gt;_&lt;seqnr&gt;.ppm&#39;, for example, &#39;MPS1300_0056.ppm&#39;. Note: the files are not in a separate directory, they will be extracted in place. However, due to the unique naming, there is no problem extracting them in one single current (destination) directory. The actual type of the image can be gray scale (.pgm) or color (.ppm), in &#39;8-bit DirectClass&#39; according to ImageMagick&#39;s &#39;identify&#39; tool. The images were cropped out of larger photographs because of irrelevant elements such as a Kodak color calibrator and non-text content such as supporting surface (table) backgrounds, seals (emblems), ribbons, etc. No effort has been made to obtain a balanced set of samples over years: the given frequencies of occurrence in archives are used. There is evidently less data in years before 1375 A.D. while some periods provides us with ample data for historical reasons (e.g, 1450 A.D.). It would have been a pity if the scarce years had determined and limited the size of this data set. Selection criteria for data reduction, whether random or systematic, would have been arbitrary. In any case, these images were used in our publications, such that the performance results of future attempts on manuscript dating can be compared with earlier results. The performances that have been reached using our algorithms are in the order of an MAE (mean average error) of 10 years. If you have any questions, please contact us: Sheng He (heshengxgd@gmail.com) Petros Samara (petros.samara@huygens.knaw.nl) Jan Burgers (jan.burgers@huygens.knaw.nl) Lambert Schomaker (L.Schomaker@ai.rug.nl) Please cite our papers if you use this data set: [1] Sheng He, Petros Samara, Jan Burgers, Lambert Schomaker. Image-based historical manuscript dating using contour and stroke fragments. Pattern Recognition(PR), Vol. 59, pp. 159-171, 2016 [2] Sheng He, Petros Samara, Jan Burgers, Lambert Schomaker. Towards style-based dating of historical documents. International Conference on Frontiers in Handwriting Recognition(ICFHR), Crete, Greece, 2014 [3] Sheng He, Petros Samara, Jan Burgers, Lambert Schomaker. Multiple-Label Guided Clustering Algorithm for Historical Document Dating and Localization IEEE Trans. on Image Processing, Vol. 25(11), Nov. 2016. http://ieeexplore.ieee.org/document/7551181/</pre> <p>Data are collected thanks to&nbsp;Dutch NWO grant project 380-50-006</p>

opencc-by-4.0Aug 2016View details →
zenodo44/100

IN00162 Cāmak चामक Charter of Pravarasena II

<p>Cāmak Charter of Pravarasena II; also called&nbsp;Chammak Plates of Pravarasena II, Chamak Plates of Pravarasena II</p>

opencc-by-4.0Sep 2019View details →
zenodo44/100

Fontenay Dataset. Original Charters From Fontenay before 1213

<p>This data set encompasses 104 images and transcriptions of digital images of original charters from the Cistercian abbey Fontenay in Burgundy (France), dating mainly from the 12th c. and until 1213.</p> <p>The original data set was created as part of the <a href="https://anr.fr/Project-ANR-12-CORP-0010">ANR ORIFLAMMS (ANR-12-CORP-0010)</a> project. Texts were transcribed in the original TEI-XML format, rendering both abbreviated and expanded forms of the original text. The alignement data was produced by merging coordinates created through the Oriflamms software.</p> <p>&nbsp;A new version was prepared in March-June 2022 as part of the research for the following paper:&nbsp;<strong>Camps</strong>, Jean-Baptiste, Chahan<strong> Vidal-Gor&egrave;ne</strong>, Dominique <strong>Stutzmann</strong>, Marguerite <strong>Vernet</strong>, and Ariane <strong>Pinche</strong>. &laquo;&nbsp;Data Diversity in Handwritten Text Recognition: Challenge or Opportunity?&nbsp;&raquo; In <em>Digital Humanities 2022. Conference Abstracts (The University of Tokyo, Japan, 25-29 July 2022)</em>, published by DH2022 Local Organizing Committee, 160‑65. Tokyo, 2022. <a href="https://dh2022.dhii.asia/dh2022bookofabsts.pdf#page=162">https://dh2022.dhii.asia/dh2022bookofabsts.pdf#page=162</a>&nbsp; and&nbsp; <a href="https://dh2022.dhii.asia/abstracts/files/CAMPS_Jean_Baptiste_Data_Diversity_in_handwritten_text_recog.html">https://dh2022.dhii.asia/abstracts/files/CAMPS_Jean_Baptiste_Data_Diversity_in_handwritten_text_recog.html</a>.</p> <p>If you use this dataset, please quote:</p> <pre> @incollection{dh2022_local_organizing_committee_data_2022, address = {Tokyo}, title = {Data {Diversity} in handwritten text recognition: challenge or opportunity?}, url = {https://dh2022.dhii.asia/dh2022bookofabsts.pdf}, language = {en}, urldate = {2022-08-02}, booktitle = {Digital {Humanities} 2022. {Conference} {Abstracts} ({The} {University} of {Tokyo}, {Japan}, 25-29 {July} 2022)}, author = {Camps, Jean-Baptiste and Vidal-Gor&egrave;ne, Chahan and Stutzmann, Dominique and Vernet, Marguerite and Pinche, Ariane}, editor = {{DH2022 Local Organizing Committee}}, year = {2022}, pages = {160--165}, } </pre> <p><br> <strong>Folders</strong><br> The present data set gathers different folders with different types of information.<br> The folder schema and file format used in the ORIFLAMMS is described in:</p> <ul> <li>Consortium Oriflamms. &laquo;&nbsp;Sp&eacute;cification du format XML-TEI pour l&rsquo;alignement texte-image. 1. Structure et convention de nommage&nbsp;&raquo;. *&Eacute;criture m&eacute;di&eacute;vale &amp; num&eacute;rique*, 11 Sept. 2016. [<a href="http://oriflamms.hypotheses.org/1442">http://oriflamms.hypotheses.org/1442</a>].</li> <li>Consortium Oriflamms. &laquo;&nbsp;Sp&eacute;cification du format XML-TEI pour l&rsquo;alignement texte-image. 2. Bonnes pratiques d&rsquo;encodage&nbsp;&raquo;. *&Eacute;criture m&eacute;di&eacute;vale &amp; num&eacute;rique*, 12 Sept. 2016. [<a href="http://oriflamms.hypotheses.org/1510">http://oriflamms.hypotheses.org/1510</a>].</li> </ul> <p><br> <strong>img</strong><br> Folder with 104 images. Provided as a separated zip file because TIF files may not be useful to all.<br> These images are scans of actual documents. All documents are preserved in the Departmental Archives of C&ocirc;te d&#39;Or (Archives d&eacute;partementales de la C&ocirc;te-d&#39;Or (https://archives.cotedor.fr/). Digitization was done by Fr&eacute;d&eacute;ric Petot.</p> <p>The source of the photograph, i.e. the shelfmark of the medieval document, is provided in the filename. &quot;FRAD021&quot; is the code for &quot; Archives d&eacute;partementales de la C&ocirc;te-d&#39;Or&quot;, then the actual shelfmark is given. They all start with &quot;15_H&quot; as &quot;15 H&quot; is the archival fonds from the Fontenay abbey. Then the sequential number has no semantic. &quot;FRAD021_15_H_257_0006&quot; means the image is the sixth reproducing documents from the archival unit &quot;15 H 257&quot; in the departmental archives of C&ocirc;te d&#39;Or (France). It is also formalized as TEI &lt;msIdentifier/&gt; element in the files of the /texts/ folder.<br> These images are also integrated in the <a href="https://bvmm.irht.cnrs.fr/">BVMM (Biblioth&egrave;que Virtuelle des Manuscrits M&eacute;di&eacute;vaux)</a> as IIIF compliant images.</p> <p><strong>texts</strong><br> Original TEI-XML edition with identifiers for all paragraphs, lines, words in the `texts/fontenay-w.xml` file, and also for characters in the `fontenay-c.xml` file.</p> <p><strong>/zones/, /img_links/</strong><br> zones described as coordinates on the images (/img/ folder) and files linking between the edition in /texts/ folder and coordinates in the /zones/ folder.</p> <p><strong>/ontologies/, /ontologies_link/, /oriflamms/</strong><br> Here, folders are present, but the data is not created (for ontologies).</p> <p><strong>/alto/</strong><br> ALTO files were created from the preexisting TEI files and combine the coordinates and the text in single files where the text is flat and rendered at a line level, with/without expansion/abbrevations for the above mentioned research: Camps, Jean-Baptiste, Chahan Vidal-Gor&egrave;ne, Dominique Stutzmann, Marguerite Vernet, et Ariane Pinche. &laquo;&nbsp;Data Diversity in Handwritten Text Recognition: Challenge or Opportunity?&nbsp;&raquo; In <em>Digital Humanities 2022. Conference Abstracts (The University of Tokyo, Japan, 25-29 July 2022)</em>, ed. DH2022 Local Organizing Committee, 160‑65. Tokyo, 2022. <a href="https://dh2022.dhii.asia/dh2022bookofabsts.pdf">https://dh2022.dhii.asia/dh2022bookofabsts.pdf</a>.</p>

opencc-by-4.0Jul 2022View details →
zenodo40/100

Maharashtra, region of ancient Vidarbha show distribution of key copper-plate charters of the Vakataka period.

<p>Maharashtra, region of ancient Vidarbha show distribution of key copper-plate charters of the Vakataka period.</p> <p>&nbsp;</p>

opencc-by-4.0Jan 2020View details →
zenodo40/100

Late Latin Charter Treebank 2 (LLCT2), version 1.2

<p>Version 1.2 of the Late Latin Charter Treebank 2 (LLCT2). Contains a number of minor corrections and&nbsp;replaces the version 1.0 published at Zenodo in 2019. Early Medieval Latin documentary texts from Italy between AD 774-897 with morphological and syntactic annotation. Latin Dependency Treebank (LDT) compatible linguistic annotation, CoNLL treebank format. Note that LLCT2 is also available open-access in the Universal Dependencies format at the <a href="https://github.com/UniversalDependencies/UD_Latin-LLCT">website</a>&nbsp;of the Universal Dependencies consortium. For a detailed description of the Late Latin Charter Treebanks, see the pre-print of the paper &#39;Late Latin Charter Treebank: contents and annotation&#39;, to be published in Corpora, 16:2 (2021), at the <a href="https://researchportal.helsinki.fi/fi/publications/late-latin-charter-treebank-contents-and-annotation">institutional repository of the University of Helsinki</a>. See also Korkiakangas, T. and Lassila, M. (2013), <a href="https://www.academia.edu/5491302/Korkiakangas_Timo_and_Lassila_Matti_Abbreviations_fragmentary_words_formulaic_language_treebanking_mediaeval_charter_material_"><em>Abbreviations, fragmentary words, formulaic language: treebanking medieval charter material</em></a>, in Mambrini, F., Passarotti, M. and Sporleder, C., <em>Proceedings of the third workshop on annotation of corpora for research in the humanities</em>, pp. 61&ndash;72, and Korkiakangas, T. and Passarotti, M. (2011), <a href="https://pdfs.semanticscholar.org/6825/a8ad70fe6e2a77540d9ff2774b8f34804fd0.pdf?_ga=2.234021022.1247020789.1580545874-419947753.1580545874"><em>Challenges in Annotating Medieval Latin Charters</em></a>, in &laquo;Journal of Language Technology and Computational Linguistics&raquo;, 26, pp. 103&ndash;114.</p>

opencc-by-4.0Jan 2020View details →
zenodo40/100

OB00061 Copper-plate charter of Harivarman

<p><a href="https://siddham.network/inscription/in00067/">IN00067</a> Copper-plate charter of Harivarman from the reign of mahārāja Budhagupta.&nbsp;The plate carries an inscription (IN00067) that registers a donation in the time of Budhagupta in year 168 of the Gupta era (equivalent to circa CE 487-88). The plate was found in Shankarpur, Sidhi District, Madhya Pradesh, India. The plate is currently stored in the Rani Durgawati Museum, Jabalpur, Madhya Pradesh. The copper plate is 24 cm x 11 cm. The inscription on the plate records that in the reign of Budhagupta, a ruler named mahārāja Gītavarman, grandson of mahārāja Vijayavarman and mahārāja Harivarman son of Rānī Svaminī and mahārāja Harivarman, donated a village named Citrapalli to a Gosvāmi brāhmaṇa. The text was written by Dūtaka Rūparāja(?), son of Nāgaśarma.</p> <p>&nbsp;</p> <p><br> The inscription was published by B. C. Jain,&nbsp;<em>Journal of the Epigraphic Society of India</em> 4 (1977): pp. 62-66 and plate facing p. 64. It was subsequently listed in&nbsp;Madan Mohan Upadhyay, <em>Inscriptions of Mahakoshal&nbsp;: Resource for the History of Central India</em> (Delhi, 2005). ISBN 81-7646496-1.</p> <p>&nbsp;</p>

opencc-by-nc-nd-4.0Mar 2017View details →
zenodo40/100

OB00609 Copper-plate charter of the Maitraka king Siladitya (plate 2).

<p>OB00609 Copper plate charter of the Maitraka king Śīlāditya I, dated year 290 [?] aśvayuja badi 10 recording a donation of villages and lands; first of two plates (OB00609 a-b) joined with a metal ring (OB00609c).</p>

opencc-by-nc-nd-4.0Apr 2017View details →
zenodo40/100

OB00609 Copper-plate charter of the Maitraka king Siladitya (plate 1).

<p>OB00609 Copper plate charter of the Maitraka king Śīlāditya I, dated year 290 [?] aśvayuja badi 10 recording a donation of villages and lands; first of two plates (OB00609 a-b) joined with a metal ring (OB00609c).</p>

opencc-by-nc-nd-4.0Apr 2017View details →
zenodo40/100

Copper-plate charter of Harivarman

<p><a href="https://siddham.network/inscription/in00067/">IN00067</a>&nbsp;Copper-plate charter of Harivarman</p> <p>&nbsp;</p>

opencc-by-4.0May 2017View details →
zenodo40/100

CHARTER WP3 Metadata

<p>CHARTER WP3 Metadata: the Excel spreadsheet contains metadata about interviews, audio and video files, and photographs taken as part the Workpackage 3 entitled "Socio-economic impacts of Arctic environmental changes on indigenous populations and local communities" in the framework of CHARTER (Drivers and Feedbacks of Changes in Arctic Terrestrial Biodiversity), between 1 August 2020 and 31 January 2025. The project was financially supported by the European Commission, grant number 869471.&nbsp;</p>

opencc-by-nc-nd-4.0Nov 2024View details →
zenodo40/100

A hybrid approach to the small unannotated corpus-based language comparison and its application to the Old East Slavic charters - Supplementary material 1 (Old East Slavic)

<h1>Old East Slavic charters (XII - XIV century)</h1> <h2>General description</h2> <p>A set of nine historical Old East Slavic legal texts from Smolensk, Polack and Novgorod from the end of the XII century to the first half of the XIV century. The source for Smolensk charters is Avanesov (1963), which contains original texts as well as their initial deciphering (not machine-readable). The source for Polack charters is Horoshkevich (2015), containing trascribed machine-readable texts that required only an additional check and preparation. The source for Novgorod charters is Napierskij (1857), carrying original texts and their initial non-machine-readable deciphering. All the texts underwent an additional preprocessing of reconstructed and contracted parts deletion in order to better represent the actual texts under consideration and exclude as many research biases as it is possible. The last step was a manual tokenisation and the joining of each text into a single string.</p> <p>The data statement is available among the downloadable files.</p> <h2>How-to</h2> <p>This section contains the tutorials that allow to use this data with the intended pipelines.</p> <h3>Corpus-based distance measurement package</h3> <p>The source code for package is available <a href="https://doi.org/10.5281/zenodo.13958502" target="_blank" rel="noopener">here</a>, the manual is available in the <a href="https://github.com/The-One-Who-Speaks-and-Depicts/corpus_distance/blob/dev/README.md" target="_blank" rel="noopener">README</a> section of the repository.</p> <p>To use this dataset for the measurement of distance between Smolensk, Polack and Novgorod lects, and their subsequent clusterisation, following steps should be completed:</p> <ol> <li>Download the <a href="https://github.com/The-One-Who-Speaks-and-Depicts/corpus_distance/blob/dev/example/Corpus_distance_tutorial.ipynb" target="_blank" rel="noopener">Jupyter notebook</a> that streamlines the package use.</li> <li>Download the dataset.</li> <li>Put the dataset into a selected folder on your computer (make sure there are no other files within this folder).</li> <li>Insert the path to the directory into <code>CONTENT_DIR</code> variable in the Jupyter notebook.</li> <li>Run the notebook, adjusting the parameters, if necessary.</li> </ol> <h2>&nbsp;</h2>

opencc-by-4.0Nov 2024View details →
zenodo40/100

A hybrid approach to the small unannotated corpus-based language comparison and its application to the Old East Slavic charters - Supplementary material 3 (Modern standard Slavic lects)

<h1>Modern standard Slavic lects (Croatian, Slovak, Slovenian)</h1> <h2>General description</h2> <p>The dataset consists of texts, written in three modern stanard Slavic lects: Croatian, Slovak, and Slovenian. The texts are parallel in order to compensate for the possible genre influences. The text is John&rsquo;s Gospel in each of the given languages.</p> <h3>Sources</h3> <p>Croatian original text is from the <a href="https://www.wordproject.org/bibles/cr/index.htm">Ivan &Scaron;arić&rsquo;s translation</a> of New Testament. Slovenian text is from the <a href="https://www.bible.com/bible/2319/JHN.1.SSP">standard Slovenian translation</a> of the New Testament. Slovak text is from the modern <a href="https://svatepismo.sk/evanjelium-podla-jana-1">Catholic translation</a> of the New Testament.</p> <p>The data statement is available among the downloadable files.</p> <h2>How-to</h2> <p>This section contains the tutorials that allow to use this data with the intended pipelines.</p> <h3>Corpus-based distance measurement package</h3> <p>The source code for package is available <a href="https://doi.org/10.5281/zenodo.13958502" target="_blank" rel="noopener">here</a>, the manual is available in the <a href="https://github.com/The-One-Who-Speaks-and-Depicts/corpus_distance/blob/dev/README.md" target="_blank" rel="noopener">README</a> section of the repository.</p> <p>To use this dataset for the measurement of distance between Slovak, Croatian and Slovenian lects, and their subsequent clusterisation, following steps should be completed:</p> <ol> <li>Download the <a href="https://github.com/The-One-Who-Speaks-and-Depicts/corpus_distance/blob/dev/example/Corpus_distance_tutorial.ipynb" target="_blank" rel="noopener">Jupyter notebook</a> that streamlines the package use.</li> <li>Download the dataset.</li> <li>Put the dataset into a selected folder on your computer (make sure there are no other files within this folder).</li> <li>Insert the path to the directory into <code>CONTENT_DIR</code> variable in the Jupyter notebook.</li> <li>Run the notebook, adjusting the parameters, if necessary.</li> </ol> <h2>&nbsp;</h2>

opencc-by-4.0Nov 2024View details →
zenodo40/100

IN00023 Dhanaidaha Charter of Kumaragupta I

<p>IN00023 Dhanaidaha Charter of Kumaragupta I</p> <p> </p> <p>Bhandarkar, Devadatta Ramakrishna, Bahadur Chand Chhabra, and Govind Swamirao Gai, <em>Inscriptions of the Early Gupta Kings</em> (New Delhi: Archaeological Survey of India, 1981): 276.</p>

opencc-by-4.0Sep 1981View details →
zenodo40/100

IN00026 Damodarpur Charter 1 (GE 124) of Kumaragupta I

<p>Bhandarkar, Devadatta Ramakrishna, Bahadur Chand Chhabra, and Govind Swamirao Gai, <em>Inscriptions of the Early Gupta Kings</em> (New Delhi: Archaeological Survey of India, 1981): 286-287.</p>

opencc-by-4.0Sep 1981View details →
zenodo40/100

IN00028 Damodarpur Charter 2 (GE 128) of Kumaragupta I

<p>Bhandarkar, Devadatta Ramakrishna, Bahadur Chand Chhabra, and Govind Swamirao Gai, <em>Inscriptions of the Early Gupta Kings</em> (New Delhi: Archaeological Survey of India, 1981): 290-291.</p>

opencc-by-4.0Sep 1981View details →
zenodo40/100

IN00035 Indor Charter of the Time of Skandagupta

<p>Bhandarkar, Devadatta Ramakrishna, Bahadur Chand Chhabra, and Govind Swamirao Gai, <em>Inscriptions of the Early Gupta Kings</em> (New Delhi: Archaeological Survey of India, 1981): 311-312.</p>

opencc-by-4.0Sep 1981View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record