Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
31
datasets available to search
ShareScore release 0.7.1
Dataset results
31 results for “Corpus Analysis”
Transcriptome analysis of canine (Canis familiaris) corpus luteum (CL) during non-pregnant late luteal phase compared with normal luteolyzing CL collected at prepartum progesterone (P4) decrease and f
GEO Series GSE98657. Canis lupus familiaris. 15 samples. Type: Expression profiling by high throughput sequencing.
miRNA analysis in myofibroblats derived from normal stomach, both antrum (A) and corpus (C) separately and gastric cancers
GEO Series GSE76218. synthetic construct; Homo sapiens. 49 samples. Type: Non-coding RNA profiling by array.
Corpus analysis of the polysemy of the verb taper 'hit' in non-standard French using Twitter data
<p>This data package contains (i) an annotated corpus sample of 10 000 tweets with the verb taper 'hit' in non-standard French (e.g. taper un coma, taper (dans) un kebab, taper une équipe 3-0) (NON-FINAL VERSION) and (ii) the Python script used to retrieve the tweets.</p>
Microarray analysis of the transcriptome in the primate corpus luteum during chorionic gonadotropin administration simulating early pregnancy.
GEO Series GSE25335. Macaca mulatta. 20 samples. Type: Expression profiling by array.
RNA Sequencing Analysis of Wild Type Caput, Corpus, and Cauda Epididymal Transcriptomes
GEO Series GSE138517. Mus musculus. 9 samples. Type: Expression profiling by high throughput sequencing.
Revised primary and secondary analysis of the DIALLS corpus
Open the record for dataset details and reuse information.
Single-cell RNA-seq analysis of subventricular zone-derived neural progenitors mobilized to demyelinated corpus callosum
GEO Series GSE144201. Mus musculus. 1 samples. Type: Expression profiling by high throughput sequencing.
Single-cell RNA-seq analysis of microglial cells in demyelinated corpus callosum
GEO Series GSE144202. Mus musculus. 2 samples. Type: Expression profiling by high throughput sequencing.
Corpus of 19th century Spanish American novels for family resemblance analysis (part of data-nh)
<p>This dataset contains different formats of a corpus of 19th century Spanish American novels which were used in a family resemblance analysis as a part of the dissertation "Genre Analysis and Corpus Design: 19th Century Spanish American Novels (1830-1910)" by Ulrike Henny-Krahmer. The texts are included as plain text files, linguistically annotated files (using TreeTagger), as text files with only noun lemmas, and as chunks of 1,000 noun tokens derived from the lemmatized texts. The texts were prepared in this way to be used with topic modeling. Because 22 of the novels still are under copyright, this dataset has restricted access. The other 234 novels are in the open domain. This dataset is part of "data-nh" (see https://github.com/cligs/data-nh), which is the whole collection of research data accompanying the above-mentioned dissertation.</p>
Bibliometric Analysis of Exhibition and Trade Fair Research: A Corpus of 69 Articles and Cluster Insights from Journal of Convention and Event Tourism
<p>This dataset comprises a bibliometric analysis of 69 articles focused on the exhibition and trade fair sector, drawn from Scopus-indexed publications within the <em>Journal of Convention and Event Tourism</em>. The dataset includes detailed information on authorship, publication trends, and the clustering of research themes, providing valuable insights into the academic landscape of this niche within the MICE (Meetings, Incentives, Conferences, and Exhibitions) industry. This resource is essential for researchers and professionals seeking to understand the development and focus areas of scholarly work in exhibitions and trade fairs</p>
Multilingual fine-grained sentiment analysis corpus
<p>A sentiment annotated corpus based on Fallout New Vegas. The corpus has the following sentiments: <em>neutral, anger, disgust, fear, happy, pained, sad, surprised</em> in the following languages: <em>English, German, Italian, Spanish and French</em>.</p> <p>Please cite the following paper: Mika Hämäläinen, Khalid Alnajjar, and Thierry Poibeau. 2022. Video Games as a Corpus: Sentiment Analysis using Fallout New Vegas Dialog. In <em>FDG’22: Proceedings of the 17th International Conference on the Foundations of Digital Games (FDG ’22)</em></p> <p> </p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.