Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
63
datasets available to search
ShareScore release 0.9.0
Dataset results
63 results for “charters”
Charter Low
First attempt at photogrammetry on 1380 Charter of City of Lincoln. Source: Objaverse 1.0 / Sketchfab
A hybrid approach to the small unannotated corpus-based language comparison and its application to the Old East Slavic charters - Supplementary material 2 (Modern East Slavic)
<h1>Modern East Slavic dialects (Belogornoje, Megra, Zialionka)</h1> <h2>General description</h2> <p>A set of modern East Slavic Belogornoje, Megra and Zialionka small territorial lects subcorpora. Megra is an autochthonous (Barannikova, 2005) Northern Russian small territorial lect (Kryuchkova and Goldin, 2011). Belogornoje is a late settlement (Barannikova, 2005) Central Russian small territorial lect (Kryuchkova and Goldin, 2011). Zialionka is an autochthonous Northern Belarusian lect, radically different from Belogornoje and Megra by most of the key isoglosses within the Eastern part of the Slavic continuum.</p> <h3>Sources</h3> <p>Both Megra and Belogornoje texts originate from the Saratov dialectological corpus (Kryuchkova and Goldin, 2011). These are manually transcribed interviews with dialect speakers, mostly on the slice-of-life, rarely touching the topic of religion, recorded during the field trips of Saratov State University from 1980 to 2019. They possess some tagging, but for the purpose of clear cross-evaluation, the experiments do not use this information. The transcription is phonemic, faithful to the dialect features, and remains untouched in the experiments.</p> <p><br>Zialionka texts are also phonemically transcribed and untouched in experiments, they come from the Polack ethnographic collection (Lobač, 2011). The main genre is folklore tales, collected by transcribing interviews with small territorial lects speakers during the field trips of Polack State University (Belarus) from 1992 to 2010 years. There are no traces of notable phonetic irregularities within the texts. Unfortunately, there is no way to reliably establish it, as there are no available original recordings.</p> <p>The data statement is available among the downloadable files.</p> <h2>How-to</h2> <p>This section contains the tutorials that allow to use this data with the intended pipelines.</p> <h3>Corpus-based distance measurement package</h3> <p>The source code for package is available <a href="https://doi.org/10.5281/zenodo.13958502" target="_blank" rel="noopener">here</a>, the manual is available in the <a href="https://github.com/The-One-Who-Speaks-and-Depicts/corpus_distance/blob/dev/README.md" target="_blank" rel="noopener">README</a> section of the repository.</p> <p>To use this dataset for the measurement of distance between Belogornoje, Megra and Zialionka lects, and their subsequent clusterisation, following steps should be completed:</p> <ol> <li>Download the <a href="https://github.com/The-One-Who-Speaks-and-Depicts/corpus_distance/blob/dev/example/Corpus_distance_tutorial.ipynb" target="_blank" rel="noopener">Jupyter notebook</a> that streamlines the package use.</li> <li>Download the dataset.</li> <li>Put the dataset into a selected folder on your computer (make sure there are no other files within this folder).</li> <li>Insert the path to the directory into <code>CONTENT_DIR</code> variable in the Jupyter notebook.</li> <li>Run the notebook, adjusting the parameters, if necessary.</li> </ol> <h2> </h2>
Advances in Distant Diplomatics. A Stylometric Approach to Medieval Charters (data and code)
<p>The materials archived in this repository accompany and support the following paper:</p> <blockquote> <p>E. Leclerq & M. Kestemont, 'Advances in Distant Diplomatics. A Stylometric Approach to Medieval Charters', in: <a href="https://riviste.unimi.it/interfaces/index">Interfaces. A Journal of Medieval European Literatures</a> [2021].</p> </blockquote> <p>The contents of this repository are the following:</p> <ul> <li><em>analysis.ipynb</em>: a Python notebook with all the code that was used for the analyses reported in the paper.</li> <li><em>CorpusCaseRogerFJeanE</em> (zipped folder): a collection of plain text files with the original charters. A primitive encoding scheme was applied to distinguish various subsections in the charters (see the Python notebook for the symbol legend).</li> <li><em>figures</em> (zipped folder): full-res versions of the automatically generated plots for the paper.</li> <li><em>embed</em> (zipped folder): interactive HTML scatterplots of the charters, colored depending on different kinds of metadata for a more intuitive exploration.</li> <li><em>hits</em> (zipped folder): HTML files containing a tabular representation of all intertexts detected between two charters (the pair of charters is identified in the filename).</li> <li><em>MetadataCaseRogerFJeanE.xlsx </em>(spreadsheet): a spreadsheet that encodes various kinds of metadata for each charter.</li> <li><em>README.md</em>: a README file.</li> </ul> <p>The original raw texts for all charters were collected from the <a href="https://www.diplomata-belgica.be/colophon_fr.html">Diplomata Belgica</a> and the <a href="http://telma.irht.cnrs.fr/outils/chartae-galliae/index/">Chartae Galliae</a> databases. In the metadata spreadsheet, the precise origin of the individual charters is given.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.