Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
277
datasets available to search
ShareScore release 0.9.0
Dataset results
277 results for “dictionary”
Modern-Botok. Custom dictionary for modern Tibetan
<p>Tsikchen.tsv is a customised dictionary to be integrated into the Tibetan tokenizer <a href="https://github.com/OpenPecha/Botok">BoTok</a>. BoTok can tokenize classical Tibetan text or traditional genres out of the box. However, since it depends on a dictionary for tokenization, it lacks capabilities for modern Tibetan, in particular, the language of modern newspapers published in the PRC or on the subcontinent. Adding this customised dictionary adds functionality for modern Tibetan to BoTok.</p> <p>The custom dictionary tsikchen was compiled from <a href="https://github.com/christiansteinert/tibetan-dictionary/tree/master/_input/dictionaries/public">Christian Steinert's collection</a> and contains the following dictionaries:</p> <ol> <li> <ol> <li>Grand Monlam Dictionary (default dictionary of Botok)</li> <li><a href="https://github.com/christiansteinert/tibetan-dictionary/blob/master/_input/dictionaries/public/07-JimValby">Jim Valby</a></li> <li><a href="https://github.com/christiansteinert/tibetan-dictionary/blob/master/_input/dictionaries/public/08-IvesWaldo">Ives Waldo</a></li> <li><a href="https://github.com/christiansteinert/tibetan-dictionary/blob/master/_input/dictionaries/public/09-DanMartin">Dan Martin</a></li> <li><a href="https://github.com/christiansteinert/tibetan-dictionary/blob/master/_input/dictionaries/public/25-tshig-mdzod-chen-mo-Tib">Tshig mdzod chen mo</a></li> <li><a href="https://github.com/christiansteinert/tibetan-dictionary/blob/master/_input/dictionaries/public/34-dung-dkar-tshig-mdzod-chen-mo-Tib">Dung dkar</a></li> <li><a href="https://github.com/christiansteinert/tibetan-dictionary/blob/master/_input/dictionaries/public/48-TibTermProject">Tibetan Terminology Project</a></li> </ol> </li> </ol> <p>The resulting dictionary was cleaned up and edited by the Divergent Discourses project to the project's requirements (removal of double entries, phraseologisms, ungrammatical entries, etc; addition of ca. 1000 personal and place names).</p> <p>For its installation, see the Divergent Discourses' <a href="https://github.com/Divergent-Discourses/modern-botok">modern-botok</a> repository on github.</p>
FIGURE 2 in A dictionary of abbreviations used in reptile descriptions
FIGURE 2. Synonyms and homonyms in traits, e.g., body length vs trunk length. There are different ways to define and measure the body or trunk length in lizards, e.g., as the distance from axilla to groin, or from collar to cloaca, or as SVL, but other definitions have been used as well. At least 7 different synonyms and homonyms have been used for trunk length (AG, AGD, AGL, BL [body length], TrunkL, TL, TRL, see Table 3). The species shown here is Dinarolacerta mosorensis (ZSM 362/1976). See text for details.
FIGURE 1 in A dictionary of abbreviations used in reptile descriptions
FIGURE 1. The use of abbreviations in reptile species descriptions has increased steadily. (A) The total number of abbreviations used in species descriptions described in a given year. (B) The number of species that use abbreviations in species descriptions since 1982. In (B) year indicates the year when the description used in our analysis was published, not the year when a species was formally described. Note that the last data points in A and B represent only 7 months of the year 2022.
国語辞書 Well thumbed dictionary 2
Source: Objaverse 1.0 / Sketchfab
Survey data on norwegian household energy use, with focus on solar PV, flexible energy use, and retrofitting in 2023 (variables dictionary)
<p>A variable dictionary (both in Norwegian and English) is added. Please note that Norwegian characters like <strong>ø</strong> may not display correctly in the webpage preview.<br>To view them properly, download the CSV file—all characters will appear correctly there.</p>
Development of Data Dictionary for neonatal intensive care unit: advancement towards a better critical care unit
Open the record for dataset details and reuse information.
RENDERING OF PHILOSOPHICAL TERMS IN EXPLANATORY DICTIONARIES OF THE RUSSIN LANGUAGE
Open the record for dataset details and reuse information.
Dictionaries [IO Bijapur 38]
<ul> <li><strong>Dictionaries.</strong></li> <li><strong>This manuscript is now IO Bijapur 38 </strong><strong>in the India Office collections.</strong></li> <li><strong>[metadata:</strong><a href="https://de.wikipedia.org/wiki/Otto_Loth"> <strong>Otto Loth, </strong></a><strong><em><a href="http://doi.org/10.5281/zenodo.3923636">A Catalogue of the Arabic Manuscripts in the Library of the India Office</a></em>, (volume 1), no. 994 here with further notations and hyperlinks]</strong>.</li> </ul> <p><a href="https://archive.org/details/catalogueofarabi01greauoft/page/276/mode/2up?view=theater"><strong>994</strong></a>.</p> <p>B 38. Size 11<sup>3/4</sup> in. by 9<sup>1/2</sup> in.; foll. 327. Seventeen lines in a page.</p> <p>A larger Dictionary of Infinitives, with explanations in <em>Persian</em>, entitled تاج المصادر ; by ABU JA’FAR Aḥmad b. ‘Alî Muḳri’ BAIHAḲÎ (nick-named Ja’farak, d.A.H. 544). See <a href="https://en.wikipedia.org/wiki/Kashf_al-Zunun">Ḥ. Kh.</a> ii. 93; Cat. Bodl. i. 234, ii. 608; and also Stewart’s Catal. 134.</p> <p>As the author states in his preface, this dictionary refers in the first place to the Koran, next to the Traditions, and lastly to ancient poetry. It is arranged in the same manner as the preceding work, and like this without any illustrative quotations.</p> <p>Boldly written, the Arabic words with vowel-points.</p> <p>Probably of the eighth century. Slightly imperfect at the end and somewhat damaged.</p> <p>The MS. was carried to <a href="https://en.wikipedia.org/wiki/Bijapur_Collection">Bîjâpûr </a>from Muḥammadâbâd (Bîdar).</p> <p>Seal of Khwâjah Jahân.</p> <p> </p> <p> </p> <p> </p> <p> </p>
Dictionaries [IO Islamic 1433]
<ul> <li><strong>Dictionaries.</strong></li> <li><strong>This manuscript is now IO Islamic 1433 </strong><strong>in the India Office collections.</strong></li> <li><strong>[metadata:</strong><a href="https://de.wikipedia.org/wiki/Otto_Loth"> <strong>Otto Loth, </strong></a><strong><em><a href="http://doi.org/10.5281/zenodo.3923636">A Catalogue of the Arabic Manuscripts in the Library of the India Office</a></em>, (volume 1), no. 1020 here with further notations and hyperlinks]</strong>.</li> </ul> <p><a href="https://archive.org/details/catalogueofarabi01greauoft/page/282/mode/2up"><strong>1020</strong></a>.</p> <p>1433. Size 10 in. by 6<sup>3/4</sup> in.; foll. 459. Twenty-one lines in a page.</p> <p>Another copy of the same work.</p> <p>Plainly written. Of the twelfth century.</p> <p>[Hastings.]</p> <p> </p>
[IO Islamic 840] Maâthir-alumarâ, The Second Edition of the Great Biographical Dictionary of the Famous Amîrs, Nawwâbs, and other Noblemen who Lived during the Reign of the Tîmûrides in India
<ul> <li><strong>Another Copy of the Maâthir-alumarâ.</strong></li> <li><strong>This manuscript is now IO Islamic 840 in the India Office collections.</strong></li> <li><strong>[metadata:</strong> <a href="https://en.wikipedia.org/wiki/Carl_Hermann_Eth%C3%A9"><strong>Hermann Ethé</strong></a><strong><em>, </em></strong><a href="http://doi.org/10.5281/zenodo.1323681"><strong><em>Catalogue of Persian Manuscripts in the Library of the India Office,</em></strong> </a><a href="http://doi.org/10.5281/zenodo.1323681"><strong>2 vols. (Oxford: India Office, 1903): volume 1,</strong></a> <strong>number 626 here with notations and hyperlinks]</strong>.</li> </ul> <p>626</p> <p>An addition to the same.</p> <p>A shorter <em>second</em> or additional volume to the preceding work, serving as supplement to the first, and containing a large number of new biographies, arranged in alphabetical order like those in the first volume. It begins with Isma’îlbeg Dûldî and concludes with Yalankûshkhân Bahâdur. No preface or khâtimah. No date. Mr. Richard Johnson received it from Mîr Muḥammad Ḥusain in Ḥaidarâbâd, A. D. 1788.</p> <p>No. 840, ff. 142, II. 21; careless Nasta’lîḳ; written, as it seems, by the same copyist who transcribed No. 622; size, 15<sup>1/8</sup> in. by 8<sup>5/8</sup> in.</p> <p> </p>
A CONSTRUÇÃO DE DICIONÁRIO BILINGUE E BIDIALETAL PARA LÍNGUA MEDZENIAKONAI (BANIWA - KORIPAKO) THE CONSTRUCTION OF A BILINGUAL AND BIDIALECTAL DICTIONARY FOR THE MEDZENIAKONAI (BANIWA KORIPAKO) LANGUAGE
Open the record for dataset details and reuse information.
Mol graph dictionary for NPLIB1 dataset
Open the record for dataset details and reuse information.
GNN graphs dictionary for molecules for NPLIB1 dataset
Open the record for dataset details and reuse information.
Online dictionary backups of Uralic languages
<p>https://sanat.csc.fi/</p>
Random Wikidata-style RDF + dictionary-based files
Open the record for dataset details and reuse information.
Magnetic Resonance Fingerprinting at 100mT with OPTIMUM: in vivo Healthy Hand Data with Dictionaries
Open the record for dataset details and reuse information.
NBG Datasets – Data Dictionary
<p>Data Dictionary for the NBG datasets</p>
Sino-Tibetan Etymological Dictionary and Thesaurus Database Software
Open the record for dataset details and reuse information.
Single-cell RNA-seq dictionary of in vivo immune responses to cytokines
GEO Series GSE202186. Mus musculus. 88 samples. Type: Expression profiling by high throughput sequencing.
Smell object in Dutch - smell in English dictionary
<p>These four csv's each contain the Dutch names of things that smell as Key (Fruits, Flowers, Nuts and Vegetables) and what they smell like as Value.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.