Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
63
datasets available to search
ShareScore release 0.9.0
Dataset results
63 results for “charters”
Late Latin Charter Treebank 1 (LLCT1), version 1.2
<p>Version 1.2 of the Late Latin Charter Treebank 1 (LLCT1). Contains a number of minor corrections, replaces the version 1.0 published at Zenodo in 2018. Early Medieval Latin documentary texts from Italy between AD 714-869 with morphological and syntactic annotation. Latin Dependency Treebank (LDT) compatible linguistic annotation, Prague style treebank format (PML). For a detailed description of the Late Latin Charter Treebanks, see the pre-print of the paper 'Late Latin Charter Treebank: contents and annotation', to be published in Corpora, 16:2 (2021), at the <a href="https://researchportal.helsinki.fi/fi/publications/late-latin-charter-treebank-contents-and-annotation">institutional repository of the University of Helsinki</a>. See also Korkiakangas, T. and Lassila, M. (2013), <a href="https://www.academia.edu/5491302/Korkiakangas_Timo_and_Lassila_Matti_Abbreviations_fragmentary_words_formulaic_language_treebanking_mediaeval_charter_material_"><em>Abbreviations, fragmentary words, formulaic language: treebanking medieval charter material</em></a>, in Mambrini, F., Passarotti, M. and Sporleder, C., <em>Proceedings of the third workshop on annotation of corpora for research in the humanities</em>, pp. 61–72, and Korkiakangas, T. and Passarotti, M. (2011), <a href="https://pdfs.semanticscholar.org/6825/a8ad70fe6e2a77540d9ff2774b8f34804fd0.pdf?_ga=2.234021022.1247020789.1580545874-419947753.1580545874"><em>Challenges in Annotating Medieval Latin Charters</em></a>, in «Journal of Language Technology and Computational Linguistics», 26, pp. 103–114.</p>
Jagjivanpur, West Bengal. Detail of the copper-plate charter of Mahendrapāla
<p>Jagjivanpur, West Bengal. Detail of the copper-plate charter of Mahendrapāla, now in the collection of the Malda Museum, Malda, West Bengal, as documented in 2007.</p>
Jagjivanpur, West Bengal. Detail of the copper-plate charter of Mahendrapāla
<p>Jagjivanpur, West Bengal. Detail of the copper-plate charter of Mahendrapāla, now in the collection of the Malda Museum, Malda, West Bengal, as documented in 2007.</p>
Charters and Records of Königsfelden Abbey and Bailiwick (1308-1662)
<p>The data has been published online as a scholarly edition: <a href="https://www.koenigsfelden.uzh.ch/">www.koenigsfelden.uzh.ch</a>.</p> <p>The charters and records have been digitized in cooperation with the <a href="https://www.ag.ch/de/bks/kultur/archiv_bibliothek/staatsarchiv/staatsarchiv.jsp">State Archives of the Aargau</a> (StAAG). All images are available in public domain. The cartularies are available via <a href="http://e-codices.ch/en/search/?aSelectedFacets=%7B%22collection_facet%22%3A%5B%22Aarau%2C+Staatsarchiv+Aargau%22%5D%7D&sQueryString=cartulary&sSearchField=fullText&sSortField=score">e-codices</a>.</p> <p>The data set has been manually prepared as a scholarly edition with information about layout and text. The data is available as <a href="https://ocr-d.de/de/gt-guidelines/trans/trPage.html">PageXML</a> and as TEI XML, prepared according to the <a href="https://tei-c.org/">TEI</a> (Text Encoding Initiative, specified by the Swiss Law Sources: <a href="https://www.ssrq-sds-fds.ch/wiki/">www.ssrq-sds-fds.ch/wiki/</a>). Further information about the scholarly edition can be found online: <a href="https://www.koenigsfelden.uzh.ch/exist/apps/ssrq/intro.html">www.koenigsfelden.uzh.ch/exist/apps/ssrq/intro.html</a>.</p> <p>The PageXML are structured in 28 collections. The identification of TEI to PageXML (and back) is given by the file names.</p> <p>The textual data is licensed under a <a href="https://creativecommons.org/licenses/by/4.0/">CC-BY 4.0 International</a> license.</p> <p>The data set is split into</p> <ul> <li>all images (except cartularies) as JPG</li> <li>all files in TEI XML</li> <li>PageXML of charters and records</li> <li>PageXML of the cartularies</li> </ul> <p>[German Abstract]</p> <p>Das Projekt «Urkunden und Akten des Klosters und Oberamts Königsfelden» will den historischen Königsfelder Urkunden- und Aktenbestand aus der Zeit bis 1662 online und als print-on-demand Buch edieren und das gesamte klösterliche Verwaltungsschriftgut digital zugänglich machen. Die Arbeiten erfolgen in Zusammenarbeit mit dem Staatsarchiv des Kantons Aargau und sind auf vier Jahre angelegt. Mit Unterstützung des Zürcher Rechtsquellenprojekts der Schweizerischen Rechtsquellenstiftung streben wir eine digitale Edition an, welche die wesentlichen Vorteile aktueller technischer Möglichkeiten nutzt, die Dokumente adäquat auszeichnet, gezielt miteinander verknüpft und für neue Zugriffsmöglichkeiten aufbereitet. Das Rückgrat des Unternehmens bildet der verhältnismässig geschlossene Bestand der überlieferten mittelalterlichen und frühneuzeitlichen Einzelblattdokumente des Klosters (STAAG U.17, 1291-1789), der - erweitert um auswärtige Stücke aus dem einstigen Kloster - in seiner historischen Entwicklung und in seinen wandelbaren Ordnungen nachvollziehbar gemacht werden soll. Aufgrund von nachträglich auf den Urkunden aufgebrachten Dorsualnotizen und Signaturen sollen historische Archivordnungen rekonstruiert sowie die späteren Abschriften dieser Urkunden nachgewiesen und verlinkt werden. So kann die Edition zugleich die Umorganisation und Umdeutung dieses Bestandes erschliessen. Wir nutzen die besondere Beweglichkeit einer digitalen Edition, um der Forschung in neuartiger Weise Entwicklungen eines Bestandes und dahinterliegende Umbrüche in der Schrift- und Administrationskultur zugänglich zu machen. Zugleich soll das Projekt Möglichkeiten digitalen Edierens im fachlichen Austausch reflektieren und vorantreiben.Mit dem historisch herausragenden Dokumenten- und Kopialbuchbestand der habsburgischen Klosterstiftung und der späteren Berner Landvogtei stellt die Edition bislang weitgehend fehlende Arbeitsgrundlagen für aktuelle Forschungsrichtungen bereit. Die Königsfelder Dokumente sind von höchstem Interesse für die Kunst-, Kultur-, Sozial- und Wirtschaftsgeschichte (vom Klosteralltag über den Klosterhaushalt bis hin zu Hofhaltung und Armenfürsorge), für die historischen Gender Studies und die Ordens- und Reformationsgeschichte. Zugleich wird unsere Edition neuen Ansätzen der Schweizer Geschichte entgegenkommen. Denn sie macht Gemengelagen zwischen unterschiedlichen Herrschaften (Habsburg, Kloster als Territorialherr, eidgenössische Orte) ebenso wie lokale Dimensionen von Staatsbildung, Territorialisierung und Konfessionalisierung fassbar. Das Projekt wird an der Universität Zürich angesiedelt, zielt auf eine enge Verkopplung von Editionsarbeit mit wissenschaftlicher Forschung und Lehre und wird Studierende mit Editionstechniken und archivalischen Gegebenheiten vertraut machen. Ausserdem wird das Projekt in Kooperation mit dem Staatsarchiv Aargau, dem Museum Aargau sowie weiteren Institutionen einen Beitrag zur Geschichtsvermittlung und Kulturgüterpflege leisten.</p> <p>Das Projekt…</p> <ul> <li>rekonstruiert und ediert einen historischen Dokumentenbestand und erschliesst weitere bislang kaum edierte Typen klösterlicher und herrschaftlicher Quellen</li> <li>unterstützt neue, über das «Werden und Wachsen der Eidgenossenschaft» hinausblickende Zugänge zur Geschichte der Schweiz</li> <li>leistet einen Beitrag zur Weiterentwicklung der offenen digitalen Edition</li> <li>wendet sich mit akademischer Lehre und über Kooperationen direkt an Studierende und ein grösseres Publikum</li> </ul>
MPS Data set with images of medieval charters for handwriting-style based dating of manuscripts
<pre>The MPS benchmark data set for handwritten manuscript dating ____________________________________________________________ This data set is collected for the Dutch NWO project: Medieval Paleographical Scale (MPS) by Petros Samara Project website: http://application02.target.rug.nl/monk/Projects/MPS/ Copyright (c) Huygensinstituut, Den Haag, 2016 University of Groningen, 2016. All rights reserved. Organisation of the data: Each .tar.gz file contains a number of NetPBM images. The format is chosen because of its simplicity. Also, there is no doubt about lossy compression in the processing chain. The file names are of the format 'MPS<year>_<seqnr>.ppm', for example, 'MPS1300_0056.ppm'. Note: the files are not in a separate directory, they will be extracted in place. However, due to the unique naming, there is no problem extracting them in one single current (destination) directory. The actual type of the image can be gray scale (.pgm) or color (.ppm), in '8-bit DirectClass' according to ImageMagick's 'identify' tool. The images were cropped out of larger photographs because of irrelevant elements such as a Kodak color calibrator and non-text content such as supporting surface (table) backgrounds, seals (emblems), ribbons, etc. No effort has been made to obtain a balanced set of samples over years: the given frequencies of occurrence in archives are used. There is evidently less data in years before 1375 A.D. while some periods provides us with ample data for historical reasons (e.g, 1450 A.D.). It would have been a pity if the scarce years had determined and limited the size of this data set. Selection criteria for data reduction, whether random or systematic, would have been arbitrary. In any case, these images were used in our publications, such that the performance results of future attempts on manuscript dating can be compared with earlier results. The performances that have been reached using our algorithms are in the order of an MAE (mean average error) of 10 years. If you have any questions, please contact us: Sheng He (heshengxgd@gmail.com) Petros Samara (petros.samara@huygens.knaw.nl) Jan Burgers (jan.burgers@huygens.knaw.nl) Lambert Schomaker (L.Schomaker@ai.rug.nl) Please cite our papers if you use this data set: [1] Sheng He, Petros Samara, Jan Burgers, Lambert Schomaker. Image-based historical manuscript dating using contour and stroke fragments. Pattern Recognition(PR), Vol. 59, pp. 159-171, 2016 [2] Sheng He, Petros Samara, Jan Burgers, Lambert Schomaker. Towards style-based dating of historical documents. International Conference on Frontiers in Handwriting Recognition(ICFHR), Crete, Greece, 2014 [3] Sheng He, Petros Samara, Jan Burgers, Lambert Schomaker. Multiple-Label Guided Clustering Algorithm for Historical Document Dating and Localization IEEE Trans. on Image Processing, Vol. 25(11), Nov. 2016. http://ieeexplore.ieee.org/document/7551181/</pre> <p>Data are collected thanks to Dutch NWO grant project 380-50-006</p>
IN00162 Cāmak चामक Charter of Pravarasena II
<p>Cāmak Charter of Pravarasena II; also called Chammak Plates of Pravarasena II, Chamak Plates of Pravarasena II</p>
Fontenay Dataset. Original Charters From Fontenay before 1213
<p>This data set encompasses 104 images and transcriptions of digital images of original charters from the Cistercian abbey Fontenay in Burgundy (France), dating mainly from the 12th c. and until 1213.</p> <p>The original data set was created as part of the <a href="https://anr.fr/Project-ANR-12-CORP-0010">ANR ORIFLAMMS (ANR-12-CORP-0010)</a> project. Texts were transcribed in the original TEI-XML format, rendering both abbreviated and expanded forms of the original text. The alignement data was produced by merging coordinates created through the Oriflamms software.</p> <p> A new version was prepared in March-June 2022 as part of the research for the following paper: <strong>Camps</strong>, Jean-Baptiste, Chahan<strong> Vidal-Gorène</strong>, Dominique <strong>Stutzmann</strong>, Marguerite <strong>Vernet</strong>, and Ariane <strong>Pinche</strong>. « Data Diversity in Handwritten Text Recognition: Challenge or Opportunity? » In <em>Digital Humanities 2022. Conference Abstracts (The University of Tokyo, Japan, 25-29 July 2022)</em>, published by DH2022 Local Organizing Committee, 160‑65. Tokyo, 2022. <a href="https://dh2022.dhii.asia/dh2022bookofabsts.pdf#page=162">https://dh2022.dhii.asia/dh2022bookofabsts.pdf#page=162</a> and <a href="https://dh2022.dhii.asia/abstracts/files/CAMPS_Jean_Baptiste_Data_Diversity_in_handwritten_text_recog.html">https://dh2022.dhii.asia/abstracts/files/CAMPS_Jean_Baptiste_Data_Diversity_in_handwritten_text_recog.html</a>.</p> <p>If you use this dataset, please quote:</p> <pre> @incollection{dh2022_local_organizing_committee_data_2022, address = {Tokyo}, title = {Data {Diversity} in handwritten text recognition: challenge or opportunity?}, url = {https://dh2022.dhii.asia/dh2022bookofabsts.pdf}, language = {en}, urldate = {2022-08-02}, booktitle = {Digital {Humanities} 2022. {Conference} {Abstracts} ({The} {University} of {Tokyo}, {Japan}, 25-29 {July} 2022)}, author = {Camps, Jean-Baptiste and Vidal-Gorène, Chahan and Stutzmann, Dominique and Vernet, Marguerite and Pinche, Ariane}, editor = {{DH2022 Local Organizing Committee}}, year = {2022}, pages = {160--165}, } </pre> <p><br> <strong>Folders</strong><br> The present data set gathers different folders with different types of information.<br> The folder schema and file format used in the ORIFLAMMS is described in:</p> <ul> <li>Consortium Oriflamms. « Spécification du format XML-TEI pour l’alignement texte-image. 1. Structure et convention de nommage ». *Écriture médiévale & numérique*, 11 Sept. 2016. [<a href="http://oriflamms.hypotheses.org/1442">http://oriflamms.hypotheses.org/1442</a>].</li> <li>Consortium Oriflamms. « Spécification du format XML-TEI pour l’alignement texte-image. 2. Bonnes pratiques d’encodage ». *Écriture médiévale & numérique*, 12 Sept. 2016. [<a href="http://oriflamms.hypotheses.org/1510">http://oriflamms.hypotheses.org/1510</a>].</li> </ul> <p><br> <strong>img</strong><br> Folder with 104 images. Provided as a separated zip file because TIF files may not be useful to all.<br> These images are scans of actual documents. All documents are preserved in the Departmental Archives of Côte d'Or (Archives départementales de la Côte-d'Or (https://archives.cotedor.fr/). Digitization was done by Frédéric Petot.</p> <p>The source of the photograph, i.e. the shelfmark of the medieval document, is provided in the filename. "FRAD021" is the code for " Archives départementales de la Côte-d'Or", then the actual shelfmark is given. They all start with "15_H" as "15 H" is the archival fonds from the Fontenay abbey. Then the sequential number has no semantic. "FRAD021_15_H_257_0006" means the image is the sixth reproducing documents from the archival unit "15 H 257" in the departmental archives of Côte d'Or (France). It is also formalized as TEI <msIdentifier/> element in the files of the /texts/ folder.<br> These images are also integrated in the <a href="https://bvmm.irht.cnrs.fr/">BVMM (Bibliothèque Virtuelle des Manuscrits Médiévaux)</a> as IIIF compliant images.</p> <p><strong>texts</strong><br> Original TEI-XML edition with identifiers for all paragraphs, lines, words in the `texts/fontenay-w.xml` file, and also for characters in the `fontenay-c.xml` file.</p> <p><strong>/zones/, /img_links/</strong><br> zones described as coordinates on the images (/img/ folder) and files linking between the edition in /texts/ folder and coordinates in the /zones/ folder.</p> <p><strong>/ontologies/, /ontologies_link/, /oriflamms/</strong><br> Here, folders are present, but the data is not created (for ontologies).</p> <p><strong>/alto/</strong><br> ALTO files were created from the preexisting TEI files and combine the coordinates and the text in single files where the text is flat and rendered at a line level, with/without expansion/abbrevations for the above mentioned research: Camps, Jean-Baptiste, Chahan Vidal-Gorène, Dominique Stutzmann, Marguerite Vernet, et Ariane Pinche. « Data Diversity in Handwritten Text Recognition: Challenge or Opportunity? » In <em>Digital Humanities 2022. Conference Abstracts (The University of Tokyo, Japan, 25-29 July 2022)</em>, ed. DH2022 Local Organizing Committee, 160‑65. Tokyo, 2022. <a href="https://dh2022.dhii.asia/dh2022bookofabsts.pdf">https://dh2022.dhii.asia/dh2022bookofabsts.pdf</a>.</p>
Maharashtra, region of ancient Vidarbha show distribution of key copper-plate charters of the Vakataka period.
<p>Maharashtra, region of ancient Vidarbha show distribution of key copper-plate charters of the Vakataka period.</p> <p> </p>
Late Latin Charter Treebank 2 (LLCT2), version 1.2
<p>Version 1.2 of the Late Latin Charter Treebank 2 (LLCT2). Contains a number of minor corrections and replaces the version 1.0 published at Zenodo in 2019. Early Medieval Latin documentary texts from Italy between AD 774-897 with morphological and syntactic annotation. Latin Dependency Treebank (LDT) compatible linguistic annotation, CoNLL treebank format. Note that LLCT2 is also available open-access in the Universal Dependencies format at the <a href="https://github.com/UniversalDependencies/UD_Latin-LLCT">website</a> of the Universal Dependencies consortium. For a detailed description of the Late Latin Charter Treebanks, see the pre-print of the paper 'Late Latin Charter Treebank: contents and annotation', to be published in Corpora, 16:2 (2021), at the <a href="https://researchportal.helsinki.fi/fi/publications/late-latin-charter-treebank-contents-and-annotation">institutional repository of the University of Helsinki</a>. See also Korkiakangas, T. and Lassila, M. (2013), <a href="https://www.academia.edu/5491302/Korkiakangas_Timo_and_Lassila_Matti_Abbreviations_fragmentary_words_formulaic_language_treebanking_mediaeval_charter_material_"><em>Abbreviations, fragmentary words, formulaic language: treebanking medieval charter material</em></a>, in Mambrini, F., Passarotti, M. and Sporleder, C., <em>Proceedings of the third workshop on annotation of corpora for research in the humanities</em>, pp. 61–72, and Korkiakangas, T. and Passarotti, M. (2011), <a href="https://pdfs.semanticscholar.org/6825/a8ad70fe6e2a77540d9ff2774b8f34804fd0.pdf?_ga=2.234021022.1247020789.1580545874-419947753.1580545874"><em>Challenges in Annotating Medieval Latin Charters</em></a>, in «Journal of Language Technology and Computational Linguistics», 26, pp. 103–114.</p>
OB00061 Copper-plate charter of Harivarman
<p><a href="https://siddham.network/inscription/in00067/">IN00067</a> Copper-plate charter of Harivarman from the reign of mahārāja Budhagupta. The plate carries an inscription (IN00067) that registers a donation in the time of Budhagupta in year 168 of the Gupta era (equivalent to circa CE 487-88). The plate was found in Shankarpur, Sidhi District, Madhya Pradesh, India. The plate is currently stored in the Rani Durgawati Museum, Jabalpur, Madhya Pradesh. The copper plate is 24 cm x 11 cm. The inscription on the plate records that in the reign of Budhagupta, a ruler named mahārāja Gītavarman, grandson of mahārāja Vijayavarman and mahārāja Harivarman son of Rānī Svaminī and mahārāja Harivarman, donated a village named Citrapalli to a Gosvāmi brāhmaṇa. The text was written by Dūtaka Rūparāja(?), son of Nāgaśarma.</p> <p> </p> <p><br> The inscription was published by B. C. Jain, <em>Journal of the Epigraphic Society of India</em> 4 (1977): pp. 62-66 and plate facing p. 64. It was subsequently listed in Madan Mohan Upadhyay, <em>Inscriptions of Mahakoshal : Resource for the History of Central India</em> (Delhi, 2005). ISBN 81-7646496-1.</p> <p> </p>
OB00609 Copper-plate charter of the Maitraka king Siladitya (plate 2).
<p>OB00609 Copper plate charter of the Maitraka king Śīlāditya I, dated year 290 [?] aśvayuja badi 10 recording a donation of villages and lands; first of two plates (OB00609 a-b) joined with a metal ring (OB00609c).</p>
OB00609 Copper-plate charter of the Maitraka king Siladitya (plate 1).
<p>OB00609 Copper plate charter of the Maitraka king Śīlāditya I, dated year 290 [?] aśvayuja badi 10 recording a donation of villages and lands; first of two plates (OB00609 a-b) joined with a metal ring (OB00609c).</p>
Copper-plate charter of Harivarman
<p><a href="https://siddham.network/inscription/in00067/">IN00067</a> Copper-plate charter of Harivarman</p> <p> </p>
CHARTER WP3 Metadata
<p>CHARTER WP3 Metadata: the Excel spreadsheet contains metadata about interviews, audio and video files, and photographs taken as part the Workpackage 3 entitled "Socio-economic impacts of Arctic environmental changes on indigenous populations and local communities" in the framework of CHARTER (Drivers and Feedbacks of Changes in Arctic Terrestrial Biodiversity), between 1 August 2020 and 31 January 2025. The project was financially supported by the European Commission, grant number 869471. </p>
A hybrid approach to the small unannotated corpus-based language comparison and its application to the Old East Slavic charters - Supplementary material 1 (Old East Slavic)
<h1>Old East Slavic charters (XII - XIV century)</h1> <h2>General description</h2> <p>A set of nine historical Old East Slavic legal texts from Smolensk, Polack and Novgorod from the end of the XII century to the first half of the XIV century. The source for Smolensk charters is Avanesov (1963), which contains original texts as well as their initial deciphering (not machine-readable). The source for Polack charters is Horoshkevich (2015), containing trascribed machine-readable texts that required only an additional check and preparation. The source for Novgorod charters is Napierskij (1857), carrying original texts and their initial non-machine-readable deciphering. All the texts underwent an additional preprocessing of reconstructed and contracted parts deletion in order to better represent the actual texts under consideration and exclude as many research biases as it is possible. The last step was a manual tokenisation and the joining of each text into a single string.</p> <p>The data statement is available among the downloadable files.</p> <h2>How-to</h2> <p>This section contains the tutorials that allow to use this data with the intended pipelines.</p> <h3>Corpus-based distance measurement package</h3> <p>The source code for package is available <a href="https://doi.org/10.5281/zenodo.13958502" target="_blank" rel="noopener">here</a>, the manual is available in the <a href="https://github.com/The-One-Who-Speaks-and-Depicts/corpus_distance/blob/dev/README.md" target="_blank" rel="noopener">README</a> section of the repository.</p> <p>To use this dataset for the measurement of distance between Smolensk, Polack and Novgorod lects, and their subsequent clusterisation, following steps should be completed:</p> <ol> <li>Download the <a href="https://github.com/The-One-Who-Speaks-and-Depicts/corpus_distance/blob/dev/example/Corpus_distance_tutorial.ipynb" target="_blank" rel="noopener">Jupyter notebook</a> that streamlines the package use.</li> <li>Download the dataset.</li> <li>Put the dataset into a selected folder on your computer (make sure there are no other files within this folder).</li> <li>Insert the path to the directory into <code>CONTENT_DIR</code> variable in the Jupyter notebook.</li> <li>Run the notebook, adjusting the parameters, if necessary.</li> </ol> <h2> </h2>
A hybrid approach to the small unannotated corpus-based language comparison and its application to the Old East Slavic charters - Supplementary material 3 (Modern standard Slavic lects)
<h1>Modern standard Slavic lects (Croatian, Slovak, Slovenian)</h1> <h2>General description</h2> <p>The dataset consists of texts, written in three modern stanard Slavic lects: Croatian, Slovak, and Slovenian. The texts are parallel in order to compensate for the possible genre influences. The text is John’s Gospel in each of the given languages.</p> <h3>Sources</h3> <p>Croatian original text is from the <a href="https://www.wordproject.org/bibles/cr/index.htm">Ivan Šarić’s translation</a> of New Testament. Slovenian text is from the <a href="https://www.bible.com/bible/2319/JHN.1.SSP">standard Slovenian translation</a> of the New Testament. Slovak text is from the modern <a href="https://svatepismo.sk/evanjelium-podla-jana-1">Catholic translation</a> of the New Testament.</p> <p>The data statement is available among the downloadable files.</p> <h2>How-to</h2> <p>This section contains the tutorials that allow to use this data with the intended pipelines.</p> <h3>Corpus-based distance measurement package</h3> <p>The source code for package is available <a href="https://doi.org/10.5281/zenodo.13958502" target="_blank" rel="noopener">here</a>, the manual is available in the <a href="https://github.com/The-One-Who-Speaks-and-Depicts/corpus_distance/blob/dev/README.md" target="_blank" rel="noopener">README</a> section of the repository.</p> <p>To use this dataset for the measurement of distance between Slovak, Croatian and Slovenian lects, and their subsequent clusterisation, following steps should be completed:</p> <ol> <li>Download the <a href="https://github.com/The-One-Who-Speaks-and-Depicts/corpus_distance/blob/dev/example/Corpus_distance_tutorial.ipynb" target="_blank" rel="noopener">Jupyter notebook</a> that streamlines the package use.</li> <li>Download the dataset.</li> <li>Put the dataset into a selected folder on your computer (make sure there are no other files within this folder).</li> <li>Insert the path to the directory into <code>CONTENT_DIR</code> variable in the Jupyter notebook.</li> <li>Run the notebook, adjusting the parameters, if necessary.</li> </ol> <h2> </h2>
IN00023 Dhanaidaha Charter of Kumaragupta I
<p>IN00023 Dhanaidaha Charter of Kumaragupta I</p> <p> </p> <p>Bhandarkar, Devadatta Ramakrishna, Bahadur Chand Chhabra, and Govind Swamirao Gai, <em>Inscriptions of the Early Gupta Kings</em> (New Delhi: Archaeological Survey of India, 1981): 276.</p>
IN00026 Damodarpur Charter 1 (GE 124) of Kumaragupta I
<p>Bhandarkar, Devadatta Ramakrishna, Bahadur Chand Chhabra, and Govind Swamirao Gai, <em>Inscriptions of the Early Gupta Kings</em> (New Delhi: Archaeological Survey of India, 1981): 286-287.</p>
IN00028 Damodarpur Charter 2 (GE 128) of Kumaragupta I
<p>Bhandarkar, Devadatta Ramakrishna, Bahadur Chand Chhabra, and Govind Swamirao Gai, <em>Inscriptions of the Early Gupta Kings</em> (New Delhi: Archaeological Survey of India, 1981): 290-291.</p>
IN00035 Indor Charter of the Time of Skandagupta
<p>Bhandarkar, Devadatta Ramakrishna, Bahadur Chand Chhabra, and Govind Swamirao Gai, <em>Inscriptions of the Early Gupta Kings</em> (New Delhi: Archaeological Survey of India, 1981): 311-312.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.