Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
2
datasets available to search
ShareScore release 0.9.0
Dataset results
2 results for “Cales”
Cariban Lexical Database (CaLeD)
<p>This dataset contains a comprehensive collection of lexical items from various languages within the Carib linguistic family. It is structured to facilitate computational historical linguistics analysis, offering detailed information on language characteristics, word forms, and cognacy judgments. The data is curated to support research in linguistic typology, historical linguistics, and related fields.</p> <h4>Data Structure</h4> <p>The dataset is presented in a TSV (Tab-Separated Values) format, ensuring easy integration with common data analysis tools. Each lexical item in the dataset is detailed with multiple linguistic attributes, including phonological transcriptions, morphological analysis, and cognacy information. The following table summarizes the fields included in the dataset:</p> <table> <tbody> <tr> <th>Field Name</th> <th>Data Type</th> <th>Description</th> </tr> </tbody> <tbody> <tr> <td>ID</td> <td>string</td> <td>Unique identifier for each dataset entry.</td> </tr> <tr> <td>ID_lang</td> <td>string</td> <td>Unique identifier for the language within the dataset.</td> </tr> <tr> <td>Glottocode</td> <td>string</td> <td>Code uniquely identifying the language in the Glottolog database.</td> </tr> <tr> <td>Glottolog_Name</td> <td>string</td> <td>Name of the language as recorded in the Glottolog database.</td> </tr> <tr> <td>ISO639P3code</td> <td>string</td> <td>ISO 639-3 code for the language.</td> </tr> <tr> <td>ID_param</td> <td>string</td> <td>Unique identifier for the linguistic parameter or concept within the dataset.</td> </tr> <tr> <td>Concepticon_ID</td> <td>integer</td> <td>Identifier for the concept in the Concepticon database.</td> </tr> <tr> <td>Concepticon_Gloss</td> <td>string</td> <td>Gloss or definition of the concept from the Concepticon database.</td> </tr> <tr> <td>Value</td> <td>string</td> <td>Value of the linguistic data point, typically a word or phrase in the language.</td> </tr> <tr> <td>Form</td> <td>string</td> <td>Phonetic or phonological transcription of the linguistic data point.</td> </tr> <tr> <td>Segments</td> <td>string</td> <td>Further phonetic or phonological breakdown of the form.</td> </tr> <tr> <td>Source</td> <td>string</td> <td>Reference to the source or citation where the data was obtained.</td> </tr> <tr> <td>Morphemes</td> <td>string</td> <td>Morphological breakdown of the form.</td> </tr> <tr> <td>SimpleCognate</td> <td>integer</td> <td>Cognacy judgment, indicating whether the form is cognate with forms of the same meaning in related languages.</td> </tr> <tr> <td>PartialCognates</td> <td>string</td> <td>Partial cognacy coding, detailing the cognacy of individual segments or morphemes.</td> </tr> </tbody> </table> <h4>Intended Use</h4> <p>This dataset is intended for researchers and linguists specializing in the Carib linguistic family. It provides valuable insights into the lexical similarities and differences across the languages within this family, supporting studies on language evolution, relationships, and structure.</p> <h4>Additional Resources</h4> <ol> <li> <p><strong>Metadata for Validation</strong>: This dataset comes with comprehensive metadata following the Frictionless Data standard, ensuring that the data structure and types are accurately described for validation purposes. This metadata aids in maintaining the integrity and usability of the data across various computational platforms and research projects.</p> </li> <li> <p><strong>CLDF Version Available</strong>: For researchers utilizing the Cross-Linguistic Data Formats (CLDF), a version of this dataset is available in CLDF specifications. This version is provided as a zipped file, facilitating easier distribution and handling.</p> </li> </ol>
FIGURES 1–4. Cales orchamoplati, female. 1 in Parasitoids of the Australian citrus whitefly, Orchamoplatus citri (Takahashi) (Hemiptera, Aleyrodidae), with description of a new Eretmocerus species (Hymenoptera, Aphelinidae)
FIGURES 1–4. Cales orchamoplati, female. 1, body; 2, antenna; 3, head, front view; 4, fore wing.
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.