Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
9
datasets available to search
ShareScore release 0.9.0
Dataset results
9 results for “Kho-Bwa”
CLDF dataset derived from Lieberherr and Bodt's "Comparative Wordlists of Kho-Bwa" from 2017
<p>Cite the source of the dataset as:</p> <blockquote> <p>Lieberherr, Ismail and Bodt, Timotheus Adrianus (2017): Sub-grouping Kho-Bwa based on shared core vocabulary. Himalayan Linguistics 16(2). 26-63. URL: https://escholarship.org/uc/item/4t27h5fg</p> </blockquote>
Data and Code Accompanying the Study on "Benefits of reflex prediction: A case study of Western Kho-Bwa"
<p>Cite the source of the dataset as:</p> <blockquote> <p>Timotheus A. Bodt and Johann-Mattis List (to appear): Benefits of reflex prediction: A case study of Western Kho-Bwa. Diachronica.</p> </blockquote>
Map of the Kho-Bwa languages
<p>This is a map of the languages of the Kho-Bwa cluster spoken in Western Arunachal Pradesh, India. The map has been used in publications such as Lieberherr and Bodt (2017: 28) and Bodt (accepted).</p> <p>Bodt, Timotheus Adrianus. accepted. <em>Proto-Western Kho-Bwa: Reconstructing the past of a small indigenous community.</em> Academia Sinica Language and Linguistics monograph series.</p> <p>Lieberherr, Ismael and Timotheus Adrianus Bodt. 2017. Sub-grouping Kho-Bwa based on shared core vocabulary. <em>Himalayan Linguistics Vol. 16(1): 2-40.</em></p>
Phylogenetic tree of the Kho-Bwa languages
<p>This is a phylogenetic tree of the Kho-Bwa languages spoken in Western Arunachal Pradesh, India. The map has been prepared using the data and methodology described in Wu, Bodt and Tresoldi (accepted). The map has also been used in Bodt (accepted).</p> <p>Wu, Mei-Shin, Timotheus A. Bodt & Tiago Tresoldi. accepted. Bayesian phylogenetics illuminate shallower relationships among Trans-Himalayan languages in the Tibet-Arunachal area. <em>Linguistics of the Tibeto-Burman Area.</em></p> <p>Bodt, Timotheus Adrianus. accepted. <em>Proto-Western Kho-Bwa: Reconstructing the past of a small indigenous community.</em> Academia Sinica Language and Linguistics monograph series.</p> <p> </p>
CLDF dataset derived from Bodt's "Lexical Cognates in Western Kho-Bwa" from 2019
<p>Cite the source of the dataset as:</p> <blockquote> <p>Bodt, Timotheus Adrianus and List, Johann-Mattis (2019): Testing the predictive strength of the comparative method: An ongoing experiment on unattested words in Western Kho-Bwa languages. Papers in Historical Phonology 4.1: 22-44.</p> </blockquote>
Prediction Experiment for Western Kho-Bwa language data: dataset
<p><strong>Prediction Experiment on Western Kho-Bwa languages</strong></p> <p><em>Timotheus A. Bodt (SOAS, London) and Johann-Mattis List (Max Planck Institute, Jena)</em></p> <p>This database includes all the sound files and the transcriptions of the prediction experiment for Western Kho-Bwa. This experiment was registered as:</p> <p>Bodt, Timotheus A., Nathan W. Hill and Johann-Mattis List. 2018. <em>Prediction experiment for missing words in Kho-Bwa language data. </em>Open Science Framework Preregistrations October 5. <a href="https://osf.io/evcbp/">https://osf.io/evcbp/</a> </p> <p>The data and code can be found on:</p> <p>Timotheus A. Bodt, Nathan W. Hill, & Johann-Mattis List. (2018, October 8). Prediction experiment for missing words in Kho-Bwa language data (Version v1.0.1). Zenodo. <a href="http://doi.org/10.5281/zenodo.1451176">http://doi.org/10.5281/zenodo.1451176</a></p> <p>A paper explaining the experiment is under review:</p> <p>Bodt, Timotheus A. and Johann-Mattis List. 2019 (under review). Testing the predictive force of the comparative method: An ongoing experiment on unattested words in Western Kho-Bwa languages. <em>Papers in Historical Phonology</em> Volume 1: 1–21.</p> <p>The results of the experiment will be presented at the International Conference on Historical Linguistics 24: 01-Jul-2019 - 05-Jul-2019, Canberra, Australia.</p> <p>The uncut sound files, cut sound files, original field notes and preliminary transcriptions have been saved as:</p> <p>Bodt, Timotheus Adrianus. (2019). <em>'Retrodiction' experiment Western Kho-Bwa languages: data [Data set]</em>. Zenodo. <a href="http://doi.org/10.5281/zenodo.2529727">http://doi.org/10.5281/zenodo.2529727</a></p> <p><strong>How to use these files?</strong></p> <ul> <li>Download the zip folder soundfiles_prediction_experiment.zip</li> <li>Extract the files in a separate folder</li> <li>Search for the required sound file(s)</li> </ul> <p>Searching sound files can best be done using the English CONCEPTS from the predictions_results.csv file. For example, searching for BACK will give all the sound files that contain the English gloss ‘back’ (including ‘backwards’, ‘back’ as body part, turn ‘back’ etc.).</p> <p>Another option is the select all the sound files of a given linguistic variety / doculect by searching for the original sound file number.</p> <p>I would advise against using a certain attested form in the predictions_results.csv file and search for that (e.g. p a ŋ + b u ‘chest’), because the cut sound files have been saved without spaces and morpheme breaks and because the actual transcriptions of the sound files may have changed after analysis, but were not updated in the name of the cut sound files.</p> <p>If you cannot find a certain sound file, then it may simply not have been recorded or not cut from the main sound file. If you are really interested, please mail me at <a href="mailto:timintibet@hotmail.com">timintibet@hotmail.com</a> and I will attempt to find it or record it.</p>
'Retrodiction' experiment Western Kho-Bwa languages: data
<p>These files contain all the data belonging to the retro-diction/pre-diction experiment in a historical-comparative linguistic study on the Western Kho-Bwa languages. This data set has all the raw as well as cut sound files of all the concepts elicited in the field work sessions in Arunachal Pradesh.</p> <p>The research was registered online as:</p> <p>Bodt, Timotheus A., Nathan W. Hill and Johann-Mattis List. 2018. <em>Prediction experiment for missing words in Kho-Bwa language data.</em> Open Science Framework Preregistrations October 5. <a href="https://osf.io/evcbp/">https://osf.io/evcbp/</a> </p> <p> </p>
Proto-Western Kho-Bwa: regionally relevant basic vocabulary elicitation list
<p>This upload contains a pdf and a word file of a 555-entry word list that was used to elicit data from eight Western Kho-Bwa varieties (Khispi, Duhumbi, Khoina, Khoitam, Jerigaon, Rahung, Rupa and Shergaon), Bangru, Brokpa and Tshangla in Western and Central Arunachal Pradesh between March 2012 and November 2018. </p> <p>The word list contains concepts that are also found in the most commonly used lists for vocabulary elicitation, but also contains regionally relevant concepts, for example, those related to flora and fauna, local livelihood practices and agricultural crops, and cultural and religious concepts. The word list is partially translated in Roman Hindi to ease elicitation with Hindi-speaking consultants, but certain concepts for which Hindi equivalents could not be found have been translated into Tshangla, the erstwhile lingua franca in some areas and a language still spoken by the older generation.</p> <p>This material is made freely available to everyone for informative or scientific purposes as long as the source (this DOI) / the collectors are properly credited. Please note that use of the material for commercial purposes <em><strong>of any kind</strong>, which includes conversion into commercial audio-visual media (documentaries etc.), storage and dissemination through sites that require registration & payment for access, or sites that rely on advertisement (including YouTube) </em>is <strong>not</strong> permitted without <strong>specific written consent</strong> from the speakers and their community, obtained through the collectors of the material. By downloading our material, you agree to these restrictions.</p> <p>This data set falls under the Attribution-NonCommercial-ShareAlike (CC BY-NC-SA) license. This license lets you remix, tweak, and build upon this work non-commercially, as long as you credit us and license your new creations under the identical terms. License Deed on <a href="https://creativecommons.org/licenses/by-nc-sa/4.0/">https://creativecommons.org/licenses/by-nc-sa/4.0/</a>. Legal Code on <a href="https://creativecommons.org/licenses/by-nc-sa/4.0/legalcode">https://creativecommons.org/licenses/by-nc-sa/4.0/legalcode</a>.</p> <p>Tim Bodt: bodttim (at) gmail (dot) com</p>
Western Kho-Bwa ethnographic background files
<p>This data set contains audio recordings, transcriptions, translations and other background files describing the ethno-linguistic history of the speakers of the eight Western Kho-Bwa varieties: Khispi, Duhumbi, Sartang (Rahung, Khoitam, Jerigaon and Khoina) and Sherdukpen (Rupa and Shergaon). These form the background data for sections 2.2. and 2.3 of the monograph "Reconstruction of Proto-Western Kho-Bwa". The details about the individual files can be found in the Table in the document "File overview.pdf".</p> <p>Bodt, Timotheus Adrianus. 2024. <em>Proto-Western Kho-Bwa: Reconstructing a communities' past through language.</em> Academia Sinica Languages and Linguistics monograph series number 67. Taipei: Academia Sinica.</p> <p><span><a href="https://www.ling.sinica.edu.tw/item/en?act=publish_book&code=view&bookID=146">LANGUAGE AND LINGUISTICS > (sinica.edu.tw)</a></span></p> <p>This material is made freely available to everyone for informative or scientific purposes as long as the source (this DOI) / the collectors are properly credited. Please note that use of the material for commercial purposes <em><strong>of any kind</strong>, which includes conversion into commercial audio-visual media (documentaries etc.), storage and dissemination through sites that require registration & payment for access, or sites that rely on advertisement (including YouTube) </em>is <strong>not</strong> permitted without <strong>specific written consent</strong> from the speakers and their community, obtained through the collector of the material. By downloading this material, you agree to these restrictions.</p> <p>This data set falls under the Attribution-NonCommercial-ShareAlike (CC BY-NC-SA) license. This license lets you remix, tweak, and build upon this work non-commercially, as long as you credit us and license your new creations under the identical terms. License Deed on <a href="https://creativecommons.org/licenses/by-nc-sa/4.0/">https://creativecommons.org/licenses/by-nc-sa/4.0/</a>. Legal Code on <a href="https://creativecommons.org/licenses/by-nc-sa/4.0/legalcode">https://creativecommons.org/licenses/by-nc-sa/4.0/legalcode</a>.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.