Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
10
datasets available to search
ShareScore release 0.9.0
Dataset results
10 results for “indigenous languages”
Convergent geographic patterns between grizzly bear population genetic structure and Indigenous language groups in coastal British Columbia
<p>Microsatellite loci calls, sex, and mean centre detection per individual (GrizzlyMicroLociMeanXY.csv) and code associated with the paper: "Convergent geographic patterns between grizzly bear population genetic structure and Indigenous language groups in coastal British Columbia". All code is from published R packages or GitHub repositories not created by the author. Code used is best described in these alternate resources. </p>
CINWA: Database of Cultivated plants and their names in the indigenous languages of South America
<p>This repository contains source data for CINWA - Database of Cultivated plants and their names in the indigenous languages of South America</p> <p><br> If you use these data please cite the database</p> <p>Aguilar Panchi, Evelyn Michelle, Saetbyul Lee, Evgenia Brodetsky, and Matthias Urban (eds.). 2022. CINWA - Database of Cultivated plants and their names in the indigenous languages of South America. Version 0.9. Available online at cinwa.org.</p> <p><br> If you would like to cite specific data entries, please also acknowledge the original source by consulting the reference that is associated with that entry. For example: Cook, Dorothy M., and Frances L. Gralow. 2001. Diccionario bilingüe koreguaje-español español-koreguaje. Santafé de Bogotá: Editorial Alberto Lleras Camargo. In: Aguilar Panchi, Evelyn Michelle, Saetbyul Lee, Evgenia Brodetsky, and Matthias Urban (eds.). 2022. CINWA - Database of Cultivated plants and their names in the indigenous languages of South America. Version 0.9. Available online at cinwa.org.</p>
CINWA: Database of Cultivated plants and their names in the indigenous languages of South America
<p><strong>This repository contains source data for CINWA - Database of Cultivated plants and their names in the indigenous languages of South America. If you use these data please cite the database. Aguilar Panchi, Evelyn Michelle, Saetbyul Lee, Evgenia Brodetsky, and Matthias Urban (eds.). 2022. CINWA - Database of Cultivated plants and their names in the indigenous languages of South America. Version 0.9. Available online at cinwa.org. If you would like to cite specific data entries, please also acknowledge the original source by consulting the reference that is associated with that entry. For example: Cook, Dorothy M., and Frances L. Gralow. 2001. Diccionario bilingüe koreguaje-español español-koreguaje. Santafé de Bogotá: Editorial Alberto Lleras Camargo. In: Aguilar Panchi, Evelyn Michelle, Saetbyul Lee, Evgenia Brodetsky, and Matthias Urban (eds.). 2022. CINWA - Database of Cultivated plants and their names in the indigenous languages of South America. Version 0.9. Available online at cinwa.org.</strong></p>
cldf-datasets/sails: South American Indigenous Language Structures (SAILS)
<p>Muysken, Pieter and Harald Hammarström and Olga Krasnoukhova and Neele Müller and Joshua Birchall and Simon van de Kerke and Loretta O'Connor and Swintha Danielsen and Rik van Gijn and George Saad. 2014. South American Indigenous Language Structures (SAILS). Jena: Max Planck Institute for the Science of Human History. (Available at <a href="https://sails.clld.org">https://sails.clld.org</a>)</p>
Timely Dictionary Development: In-person and Virtual Rapid Word Collection for Endangered Indigenous Languages
<p>Timely Dictionary Development: In-person and virtual Rapid Word Collection for endangered Indigenous Languages</p> <p>Dorothea Hoffmann, Wilhelm Meya, Abbie Hantgan-Sonko & Elliot Thornton</p> <p>Presented 5 October 2022 at the Berlin-Brandenburg Academy of Sciences and Humanities Where Do We Need to Go From Here? Language Documentation and Archiving in the International Decade of Indigenous Languages</p>
Music Data, Archiving for Community Use and Future Directions Through the Decade of Indigenous Languages
<p>Music Data, Archiving for Community Use and Future Directions Through the Decade of Indigenous Languages</p> <p>Linda Barwick</p> <p>Presented 5 October 2022 at the international conference "Where Do We Need to Go From Here?" Language Documentation and Archiving in the International Decade of Indigenous Languages</p>
Projected speaker numbers and dormancy risks of Canada's Indigenous languages
<p>Data and code for the paper "Projected speaker numbers and dormancy risks of Canada’s Indigenous languages".</p> <p>Contain speaker numbers by age and language (Indigenous mother tongue, unique responses). </p> <p>There is one file per year (2001, 2006, 2011, 2016, 2021).</p> <p>Data were provided by Statistics Canada.</p> <p>These include: </p> <p>- indigenousmothertongue2001.csv<br>- indigenousmothertongue2006.csv<br>- indigenousmothertongue2011.csv<br>- indigenousmothertongue2016.csv<br>- indigenousmothertongue2021.csv</p> <p>Additionally, the file 'coordinates.xlsx' contains the geographic coordinates necessary for Fig. 1. Information comes from Ethnologue with modifications.</p> <p>Also included is the life table information produced by World Population Prospects 2024 (wpp 2024 files) available at https://population.un.org/wpp/Download/Standard/Mortality/. These are provided here for convenience as well as to prevent updates by the WPP. </p> <p>Also contains the whole R code to produce the results described in the paper (RevisedScript_ProjectCanIndigLangs_Final.R).</p>
cldf-datasets/sails: South American Indigenous Language Structures (SAILS)
<p><strong>Muysken, Pieter and Harald Hammarström and Olga Krasnoukhova and Neele Müller and Joshua Birchall and Simon van de Kerke and Loretta O'Connor and Swintha Danielsen and Rik van Gijn and George Saad. 2014. South American Indigenous Language Structures (SAILS). Jena: Max Planck Institute for the Science of Human History. (Available at <a href="https://sails.clld.org/">https://sails.clld.org</a>).</strong></p>
Comments on five articles in The Conversation (Australia) on Indigenous language revitalization
<p><em>This dataset consists of comments collected from five articles on the topic of Indigenous language revitalization, published on </em><a href="https://theconversation.com/au"><em>theconversation.com/au</em></a><em> between August 2014 and July 2020. </em><em>The titles, publication date, and URL of each article are listed below. </em></p> <p> </p> <p><em>The comments presented in this dataset include only first tier comments, i.e., those that are made directly on the article, but excluding comments on those comments and subsequent discussion threads. </em></p> <p> </p> <p><em>Comments were made in a public forum using both real names and pseudonyms; these have been retained in the dataset. Comments have not been edited in any way, though numbers were added to help distinguish them. </em></p> <p> </p> <p><em>This data was collated by Dr. Gerald Roche, Senior Research Fellow in the Department of Politics, Media and Philosophy at La Trobe University in October 2020.</em></p>
Data to accompany dissertation: Geographic, Cultural, and Ecological Correlations with Indigenous Language Vitality in North America
<p>Text files:</p> <ul> <li>readme_general.txt contains a brief description of files included.</li> <li>readme_modeldata.txt contains a metadata description of model_data.csv.</li> <li>readme_languageslandNorthAmerica.txt contains a metadata description of Languages_land_NorthAmerica.csv.</li> <li>readme_languagerevitalizationdatabase.txt contains a metadata description of Language_revitalization_database.csv.</li> </ul> <p>CSV files:</p> <ul> <li>Languages_land_NorthAmerica.csv is a version of the Languages of Government-Recognized Native Land Areas in the Continental United States database. It includes data from the US Census 2017 TIGER/Line AIANNH shapefile with one row per Native land area and additional columns for associated information that was coded and calculated for this dissertation as discussed in Section 3.3.1.</li> <li>Language_revitalization_database.csv is the Language Revitalization Database. It contains the master language list used for this dissertation and columns created while coding data for the language revitalization variable, as discussed in Section 3.3.2.</li> <li>model_data.csv contains data for all variables used in the analysis and is the .csv file needed to run LanguageVitalityModels.R.</li> </ul> <p>R scripts:</p> <ul> <li>LanguageVitalityModels.R is the R script for the main part of the dissertation analysis.</li> </ul>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.