Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
5
datasets available to search
ShareScore release 0.9.0
Dataset results
5 results for “subculture”
Salvaging the Internet Hate Machine: Using the discourse of extremist online subcultures to identify emergent extreme speech
<p>This dataset accompanies a paper submitted to the WebSci 20 conference. In this paper, we present a lexicon of 'extreme speech' that may be used to detect hate speech and extreme speech on online platforms. We outline a cross-disciplinary research protocol through which this lexicon is initially extracted from a corpus of 3,335,265 posts from 4chan's /pol/ sub-forum using a hybrid method comprising word2vec modeling and subsequent snowballing of nearest neighbours of a small initial expert seed list of extreme language. The choice of corpus is significant, as 4chan is a space of rapid language innovation and obscure extreme vernacular, complicating generalised approaches. Our lexicon detects significantly more extreme posts within a corpus from a more mainstream platform (Reddit) than another popular lexicon, Hatebase, with similar accuracy. Our lexicon and the method of its creation thus provide a contribution to the study of the toxicity of online subcultures similar to 4chan, as well as more mainstream platforms. As we demonstrate, the lexicon allows for more effective detecting of extreme speech in these spaces. This method and the lexicon have further been made available through an open-source web tool for the study of online social platforms, 4CAT. The computational methods and lexicon on offer here can thus be used by a wide academic audience, fostering interdisciplinary approaches to the study of online hate and extreme speech. </p> <p>The dataset comprises the following items:</p> <ul> <li>The 4chan corpus from which the extreme speech lexicon was generated (posts from /pol/, 1 October 2019 - 1 November 2019)</li> <li>The Reddit corpus used to verify and test the lexicon (posts from the_donald, theredpill, politics and chapotraphouse, 1 October 2019 - 1 November 2019)</li> <li>The word2vec model from which the extreme speech lexicon was generated</li> <li>The extreme speech lexicon that was generated</li> </ul>
Transcriptome sequencing of the callus in subculture 21d of YIL25 versus Teqing, the complementary transgenic line (pCPL) versusTeqing , the overexpressiong transgenic line (pOE) versusTeqing .
GEO Series GSE142094. Oryza sativa. 18 samples. Type: Expression profiling by high throughput sequencing.
Weimar (East Germany), house project for art, colourful variety and subculture at Gerberstraße 3
<u>Source</u>: Flickr <br><u>4DCity URL</u>: <a href="https://4dcity.org/imgupload/1653149661.6439.jpg">https://4dcity.org/imgupload/1653149661.6439.jpg</a> <br><u>Original Image URL</u>: <a href="https://live.staticflickr.com/65535/51549087193_c43a051144_m.jpg">https://live.staticflickr.com/65535/51549087193_c43a051144_m.jpg</a> <br><br><u>Image-Metadata:</u><br>Filename: 1653149661.6439.jpg<br>Image Dimensions: 240x160<br>Megapixels: 0.04 MP<br>Filesize: 26.77 KB<br><br>Copyright: Ralf Steinberger<br>ExifOffset: 86<br>Artist: Ralf Steinberger
Genome-Wide Transcriptome Analysis Reveals Degeneration Progress of Subculture Cordyceps Militaris
GEO Series GSE100834. Cordyceps militaris. 6 samples. Type: Expression profiling by high throughput sequencing.
In-Depth Characterization of Cell Subculture Process based on Multi-Omics
GEO Series GSE215840. Chlorocebus aethiops. 30 samples. Type: Expression profiling by high throughput sequencing.
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.