Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
63
datasets available to search
ShareScore release 0.9.0
Dataset results
63 results for “verbs”
BOLD Verb Generation
Open the record for dataset details and reuse information.
Visual Dictionary of Tibetan Verb Valency: Data
<p>This repository contains a JSON version of the data powering the <a href="https://bit.ly/VisualDictionary-TibetanValency">Visual Dictionary of Tibetan Verb Valency</a> together with its documentation. The structure of the data is explained in the documentation section on 'Dictionary data'.</p> <p>The Visual Dictionary of Tibetan Verb Valency was produced as part of the UKRI-funded project <a href="https://gtr.ukri.org/projects?ref=AH%2FP004644%2F1">Lexicography in Motion a History of the Tibetan Verb</a> (LIM) at SOAS.</p>
Between syntax and morphology: German noun+verb units (Glossa)
<p><strong>This dataset accompanies a paper to be published in Glossa. Under the present DOI, all data generated for this research as well as all scripts used are stored. The paper itself is CC-licensed, refer to glossa-journal.org.</strong></p><p><strong>Abstract</strong></p><p>We show that graphemic variation—at least in some writing systems—can be analysed in terms of grammatical variation given a usage-based probabilistic view of the grammar-graphemics interface. Concretely, we examine a type of noun+verb unit in German, which can be written as one word or two. We argue that the variation in writing is rooted in the units' ambiguous status in between morphology (one word) and syntax (two words). The major influencing factors are shown to be the semantic relation between the noun and the verb (argument or oblique relation) and the morphosyntactic context. In prototypically nominal contexts, a re-interpretation of the unit as a noun+noun compound is facilitated, which favours spelling as one word, while in prototypically verbal contexts, a syntactic realisation and consequently spelling as two words is preferred. We report the results of two large-scale corpus studies and a controlled production experiment to corroborate our analysis.</p>
The Audio and Video Files of Handbook of Japanese Basic Verbs
<p>"KIHONDOUSHI HANDBOOK (Handbook of Japanese Basic Verbs) [基本動詞ハンドブッ ク] is a digital reference resource developed by the members of three collaborative research projects funded by the National Institute for Japanese Language and Linguistics (NINJAL), Tokyo, Japan during Oct. 2009 to March 2022. The reference work was developed with an aim to deepen the understanding of highly polysemous basic verbs in Japanese by way of providing detailed description of syntax and semantics of multiple meanings of the verb in question. To foster the ease of understanding, each meaning of a polysemous basic verb is illustrated with a couple of illustrative examples with audio and a few meanings, which are deemed to be difficult to understand, are illustrated with the help of audio-visual animations. The handbook consists of 190 headwords and includes 2,154 meanings, 12, 492 illustrative examples with audio and 521 audio-visual animations. On this site the audio and video files of the handbook is made available for researchers, teachers, and learners of Japanese language for the purpose of research on and education of Japanese polysemous basic verbs. Other data such as headword data will be available from the NINJAL repository site very soon.</p>
Kobalt: Extension Corpus and Annotation Guidelines for Verb Classification and Dependency Adjustments
<p>Kobalt (Zinsmeister et al. 2012) is a task-based corpus of essays written by learners and native speakers of German. This repository contains data that was not included in the original corpus and new layers of annotation to the original and the extended corpus, specifically morphological and syntactic classification of verbs and corrections and changes to dependency parses. Please refer to the annotation guidelines included in this repository for further information.<br> </p>
Inchoative -gyettiđ/-škyettiđ verbs in Inarilappisches Wörterbuch
<p>Lexical data extracted from (a scanned version of): InLpWb = Inarilappisches Wörterbuch 1–3. Compiled by Erkki Itkonen. Société Finno-Ougrienne, Helsinki 1986–1989. The table contains all the verbs with a suffix -gyettiđ/-guáttiđ or -škyettiđ/-škuáttiđ in InLpWb. Derivational bases and ultimate root lexemes are given when possible. A semantic categorization of inchoative/momentative/translative is provided for the -gyettiđ/-guáttiđ verbs.</p> <p>This is a data set supplement of: Kuokkala, J. (2019). Saami <em>-(š)goahti</em> inchoatives: their variation, history, and suggested cognates in Veps and Mordvin. <em>Suomalais-Ugrilaisen Seuran Aikakauskirja</em> 97, 153–181. <a href="https://doi.org/10.33340/susa.75718">https://doi.org/10.33340/susa.75718</a></p>
Italian Verb Lexicon for Sentiment Inference
<p><strong>Italian Verb Lexicon for Sentiment Inference</strong></p> <p><strong>Theory:</strong></p> <p>For a description of the theory behind the specifications of the corpus, please read the attached paper. </p> <p><br> <strong>Example of json entry:</strong></p> <p>{"verb": "soddisfare", "frames": [{"fillers": ["Subj", "DirObj/IndObj"], "polarity": "POS", "effects": [["DirObj/IndObj", "pos"]], "expectations": [], "examples": ["L'offerta ha soddisfatto i clienti.", "Soddisfare al pubblico."], "remarks": [], "relations": [["Subj", "DirObj/IndObj", "pro"]]}]}</p> <p><strong>Description: </strong><br> The verb "soddisfare" has 2 frames, a subject followed by a direct or indirect object. The verb polarity is positive. There is an positive effect on the direct (indirect) object. No expectations. There is a in favour (pro) relation from the subject to the direct (indirect) object. Two example sentenes are given.</p> <p><strong>Synonyms:</strong><br> Some entries are references to synonym verbs with identical frames:</p> <p>{"verb": "consacrare", "germanTranslation": "widmen", "frameReference": "dedicare", "examples": ["Consacrare tempo alle sue passioni"]}</p> <p>Here, "cosacrare" and "dedicare" are assumed synonyms with the same syntactic frames.</p> <p><strong>Used Tags:</strong></p> <p>A few explanations on the tags used in the verb specifications sheets.</p> <p>Subj = subject</p> <p>DirObj = direct object, as in "Il professore legge __il giornale__".</p> <p>IndObj = indirect object, as in "Permettere qualcosa __a qualcuno__".</p> <p>RefObj = reflexive object (pronoun), as in "La squadra avversaria __si__ è arrabbiata moltissimo". </p> <p>PrepObj[prep] = prepositional phrase; the preposition is specified in the square brackets. If more than one preposition can occur,<br> no specification is given.</p> <p>SubCl = a subordinate clause, usually introduced by "che" or "di" such as in "Ha detto __di andarsene__", <br> "Ha detto __che tutto è andato bene__".</p> <p>mod = any type of modifier, mostly adverbs, e.g. "Se ne è andato __subito__".</p> <p><br> *, e.g. mod* = indicates optionality</p> <p><br> </p> <p> </p>
Tables of inchoative verbs in the Saami languages
<p>The tables in the .xlsx file present the occurrence of various types of derived inchoative verbs by tabulating the attested derivatives of each base verb in each Saami language. Base verbs are given in their Proto-Saami form and approximate meanings in Finnish. For more information about sources etc., see the paper which this is an appendix to:</p><p>Kuokkala, Juha. 2023. Inchoative verbs in Saami: Derivational types and their variation. In: <i>Saami Linguistics in Uppsala: The 2019 SAALS 4 Symposium</i>. Studia Uralica Upsaliensia 42.</p><p>Changelog:</p><ul><li>Version 1.1 - Added missing North Saami <i>balˈlát</i></li></ul>
Data and results for "A corpus-based study to triangulating experimental evidence regarding verb-noun association for action verbs"
<p>This repository provides spreadsheets containing the results of corpus-based and experimental studies for my undergraduate thesis titled "A corpus-based study to triangulating experimental evidence regarding verb-noun association for action verbs" (supervised by Gede Primahadi Wijaya Rajeg, PhD [main] and Ketut Santi Indriani, M.Hum. [associate]) in the Bachelor of English Literature (BoEL) program, Faculty of Humanities, Udayana University. The thesis explores convergences/divergences between different methods and data types for a set of verb-noun collocations for several action verbs and their synonyms. The description of the dataset is as follows:</p> <ol> <li>"data-raw": A raw dataset containing the results of an experiment conducted using Gorilla Experiment Builder. This consists of responses regarding verb-noun collocation co-occurrences from 17 participants into one. Link to Gorilla Experiment: (https://app.gorilla.sc/openmaterials/622948).</li> <li>"Corpus Analysis Results": A compiled data containing search results of frequencies found in the Corpus of Contemporary American English (COCA). The frequencies were compiled into tables for the five main verbs showing the number of co-occurrences of specific verb-noun collocations.</li> <li>"Experiment Results (1)": A compiled data containing the experiment results calculated as a total, showing the number of co-occurrences of specific verb-noun collocations across five verbs from the participant responses.</li> </ol> <p>The thesis is part of the pedagogical outcome of the <a title="CompLexico" href="https://www.cirhss.org/complexico/" target="_blank" rel="noopener"><em>CompLexico</em></a> research group at <a title="CIRHSS" href="https://www.cirhss.org/" target="_blank" rel="noopener"><em>CIRHSS</em></a>, and the Psycholinguistics course I took with I Made Sena Darmasetiyawan, PhD at BoEL, both in the Faculty of Humanities, Udayana University.</p>
Dataset: Verb Technology Company, Inc. (VERB) Stock Performance
This dataset provides historical stock market performance data for specific companies. It enables users to analyze and understand the past trends and fluctuations in stock prices over time. This information can be utilized for various purposes such as investment analysis, financial research, and market trend forecasting.
The usage of phrasal verbs (emu) ye den, (emu) ye duru and (emu) ye hare in Akan Twi Asante
<p>A project done as part of the "Urbane Feldforschung" seminar in Humboldt University of Berlin in 2018. This is a project on Akan Twi Asante, a language spoken in Ghana. The subject is the usage of phrasal verbs (emu) ye den, (emu) ye duru and (emu) ye hare in the Asante dialect of Akan Twi.</p> <p>Collected data includes 6 glossed sentences in Akan on the topic mentioned above and 56 words from the Swadesh List. <br> <br> <br> <br> </p>
The usage of phrasal verbs (emu) ye den, (emu) ye duru and (emu) ye hare in Akan Twi Asante
<p>A project done as part of the "Urbane Feldforschung" seminar in Humboldt University of Berlin in 2018. This is a project on Akan Twi Asante, a language spoken in Ghana. The subject is the usage of phrasal verbs (emu) ye den, (emu) ye duru and (emu) ye hare in the Asante dialect of Akan Twi.</p> <p>Collected data includes 6 glossed sentences in Akan on the topic mentioned above and 56 words from the Swadesh List. <br> </p>
ERP evidence of embodiment of action-verbs at lexical stages in L1 and L2
<p>EEG data for the article : Britz J., Collaud E., Jost L., Sato S., Bugnon A., Mouthon M. and Annoni JM. ERP evidence of embodiment of action-verbs at lexical stages in L1 and L2. Brain sciences 2024</p> <p>The data used in the study were organized using the Brain Imaging Data Structure (BIDS) (Gorgolewski, K., Auer, T., Calhoun, V. et al., 2016) with the extension for EEG data (Pernet, C.R., Appelhoff, S., Gorgolewski, K.J. et al., 2019).</p> <p> </p> <p>.....</p>
Extension of the Action Verb Corpus (AVCext)
<p>Extension of the Action Verb Corpus (AVCext)</p> <p>The extension to the <a href="https://zenodo.org/record/5140014#.YQAkEjqxWEI">Action Verb Corpus</a> consists of 41 recordings conducted by 2 users experienced with the system performing the same three actions as in AVC — <strong>take</strong> (208 instances), <strong>put</strong> (208 instances), and <strong>push</strong> (91 instances). The actions were performed without any instructions. The focus of the extension is to facilitate action recognition. In comparison to AVC, no speech-related information is annotated in the extension dataset. The other ELAN annotations available with AVC are also available for the extension dataset. Additionally, the <strong>actions are annotated in two degrees of granularity. Coarse labels are: take, put and push. Fine labels split the motion into more granular motion primitives: reach, grab, moveObject, and place.</strong> These annotations are available as eaf (ELAN) files, csv files and as two separate columns in the Merged files. Details about the collected data can be found in <a href="https://www.researchgate.net/profile/Stephanie-Gross-schreitter/publication/326640788_Extension_of_the_Action_Verb_Corpus_for_Supervised_Learning/links/5be2a5c3299bf1124fc09f40/Extension-of-the-Action-Verb-Corpus-for-Supervised-Learning.pdf">Matthias Hirschmanner, Stephanie Gross, Brigitte Krenn, Friedrich Neubarth, Martin Trapp and Markus Vincze: Extension of the Action Verb Corpus for Supervised Learning. ARW 2018.</a></p> <p>The dataset consists of the following information:</p> <ul> <li>the merged output of the hand and object trackers (AVCExtension_Merged.zip, one csv file per episode/recording) <ul> <li>HandID: 0 right, 1 left</li> <li>FingerID: 0 thumb, 1 index, 2 middle, 3 ring, 4 pinky</li> <li>BoneID: 0 metacarpal, 1 proximal, 2 intermediate, 3 distal</li> </ul> </li> <li>the output of the object trackers, including the object poses and their reliability estimate calculated by the object tracker, whether an object is touched by or is in the hand of the instructor and whether the object touches the table, for the coordinate system applied see picture "coord_system.png (AVCExtension_Objects.zip, one csv file per episode/recording),</li> <li>the videos from Leap Motion showing the hand movements and objects<br> (AVCExtension_video_libm.zip, one avi file per episode/recording),</li> <li>animation of the merged hand and object tracking<br> (AVCExtension_video_schematic.zip, one avi file per episode/recording),</li> <li>the following list of annotations synchronized with the real-time animation of the hand and object tracking (available as <a href="https://tla.mpi.nl/tools/tla-tools/elan/">ELAN</a> files AVCExtension_Annotations_eaf.zip, and csv files AVCExtension_Annotations_csv.zip, one file per episode/recording) <ul> <li>information which object is currently moved, and where it is moved to (automatically annotated),</li> <li>information whether a hand touches a particular object (manually annotated),</li> <li>information whether a particular object touches the ground/table (automatically annotated),</li> <li>coarse-grained annotation: take, put, push (manually annotated),</li> <li>fine-grained annotation: reach, grab, moveObject, and place (manually annotated),</li> <li>position of the objects in the scene (automatically calculated from output of object tracker)</li> </ul> </li> </ul> <p>Acknowledgments</p> <p>Corpus creation and annotation was supported by the <a href="http://www.wwtf.at/">WWTF</a> project <a href="http://ralli.ofai.at/home.html"> RALLI</a>. The dataset was recorded at <a href="https://www.acin.tuwien.ac.at">ACIN, TUW</a>.</p>
*-koatē/-škoatē inchoative verbs in Skolt and Kola Saami dialects
<p>The file contains two sets of data concerning *<em>-koatē/-škoatē</em> inchoative verbs in Skolt and Kola Saami dialects. The first one is a schematic representation of the verbs found in KKLS [1] (alphabetical span of A–K) with a focus on the derivational variants *<em>-koatē</em> vs. *<em>-škoatē</em>. Two versions are presented on separate sheets, one with a single alphabetical listing and another with grouping according to the (original) syllable count of the verb base. The columns P...V code the variants occurring in different dialects (codes: k = *-koatē, s = *-škoatē, ks = both).</p> <p>[1] KKLS = T. I. Itkonen 1958: <em>Koltan- ja kuolanlapin sanakirja. Wörterbuch des Kolta- und Kolalappischen</em>. I–II. Helsinki: Société Finno-Ougrienne.</p> <p>The second data set contains the *<em>-koatē/-škoatē</em> verb lexemes found in Szabó 1968 and 1987 [2, 3]. The data is grouped according to (original) syllable count, suffix variant (*<em>-koatē</em> vs. *<em>-škoatē</em>) and language area (Kildin vs. Ter [Turja]). Only one occurrence of each derivative lexeme is quoted.</p> <p>[2] Szabó, László 1968: <em>Kolalappische Volksdichtung (Texte aus den Dialekten in Kildin und Ter)</em>. Zweiter Teil nebst grammatischen Aufzeichnungen. Göttingen: Vandenhoeck & Ruprecht.</p> <p>[3] Szabó, László 1987: The use of the inchoative in Kola-Sami sentences. – <em>Nordlyd </em>13: 70–103.</p> <p>This is a data set supplement of: Kuokkala, J. (2019). Saami -<em>(š)goahti</em> inchoatives: their variation, history, and suggested cognates in Veps and Mordvin. <em>Suomalais-Ugrilaisen Seuran Aikakauskirja</em> 97, 153–181. <a href="https://doi.org/10.33340/susa.75718">https://doi.org/10.33340/susa.75718</a></p>
Supplementary materials for "Spanish lower and upper bounded change of state verbs: Focusing on transitive experiencer object verbs", published in Linguistics: An Interdisciplinary Journal of the Language Sciences
<p>Supplementary files are as follows:</p> <ul> <li><strong>R-code_Scalarity.Rmd</strong> contains the R-code in a markdown version of the statistical analysis of the study.</li> <li><strong>R-code_Scalarity.html</strong> displays the R-code_Scalarity.Rmd in an html format.</li> <li><strong>Results_Scalarity.csv</strong> contains the data of the study.</li> </ul> <p>For more details on the description of the data, see the publication: "Spanish lower and upper bounded change of state verbs: Focusing on transitive experiencer object verbs", Linguistics.</p>
The neural correlates of embodied L2 learning Does embodied L2 verb learning affect representation and retention?
<p>We investigated how naturalistic actions in a highly immersive, multimodal, interactive 3D virtual reality (VR) environment may enhance word encoding by recording EEG in a pre/post-test learning paradigm. While behavior data has shown that coupling word encoding with gestures congruent with word meaning enhances learning, the neural underpinnings of this effect have yet to be elucidated. We coupled EEG recording with VR to examine whether “embodied learning” improves learning and creates linguistic representations that produce greater motor resonance. Participants learned action verbs in an L2 in two different conditions: Specific action (observing and performing congruent actions on virtual objects) and Pointing (observing actions and pointing to virtual objects). Pre and post-training participants performed a Match-mismatch task as we measured EEG (variation in the N400 response as a function of match between observed actions and auditory verbs) and a Passive listening task while we measured motor activation (mu (8-13 Hz) and beta band (13-30Hz) desynchronization during auditory verb processing) during verb processing. Contrary to our expectations, post-training results revealed neither semantic nor motor effects in either group when considered independently of learning success. Behavioral results showed both groups learned the verbs, but also a great deal of variability in learning success. When considering performance, Low performance learners showed no semantic effect and High performance learners exhibited an N400 effect for Mismatch vs Match trails post-training, independent of the type of learning. Taken as a whole, our results suggest that embodied processes can play an important role in L2 learning.</p>
Dataset and documented R code for "Nouns and verbs in the speech signal"
<p>The files available constitute supplementary material to the following article:</p> <p>Lohmann, Arne. Nouns and verbs in the speech signal: Are there phonetic correlates of grammatical category? <em>Linguistics</em> - <em>An Interdisciplinary Journal of the Language Sciences</em>.</p> <p>The article is to be published online in 2020, and in 2021 in the print version of the journal.</p>
An xml collection of example sentences for Lhasa Tibetan verbs
<p>This is an xml collection of example sentences for Lhasa Tibetan verbs</p>
Appendix_verb_scores_UD
<p>A dataset that plots verb scores (1 to 3) in main and adverbial clauses in the sample languages covered by UD.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.