Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
6
datasets available to search
ShareScore release 0.7.1
Dataset results
6 results for “Weibo”
Data for "Demographic inequalities in digital spaces in China: The case of Weibo"
<p>These data underlie the results and figures used in the article "Demographic inequalities in digital spaces in China: The case of Weibo" (https://doi.org/10.36190/2023.01). This research was presented at the ICWSM workshop "Data for the wellbeing of the most vulnerable" on June 5, 2023.</p><p>The corresponding workflow can be found in the linked repository.</p><p>`README.md` provides more details.</p>
Dataset: Weibo Corporation (WB) Stock Performance
This dataset provides historical stock market performance data for specific companies. It enables users to analyze and understand the past trends and fluctuations in stock prices over time. This information can be utilized for various purposes such as investment analysis, financial research, and market trend forecasting.
Wbbyyr: FastText language models for Mandarin Chinese, trained on 14m Sina Weibo posts for each year in 2012-2018 (Fold 1 of 10)
<p>Wbbyyr: FastText language models for Mandarin Chinese, trained on 14,440,000 Sina Weibo posts for each year in 2012-2018.</p> <p>The 14,440,000 posts from each year are split into 10 folds. Due to Zenodo size limit, this dataset contains only the first fold from each year.</p> <p>Each model is trained for 20 iterations. Each vector is 300 dimensions long.</p>
Weibo-Covid-19
<p>Weibo is a popular Chinese-language messaging platform. The dataset contains all messages that included one of the following terms in Chinese: “vaccine,” “vaccination,” “adenovirus vector,” “inactivation,” “clinical trail, ” “Phase III trail,” “immune,” “antibody,” “mutant virus,” “herd immunity,” “novel<br>coronavirus,” “corona virus,” “Covid,” “novel corona pneumonia,” “Wuhan’s unknown pneumonia,” “pneumonia of unknown<br>cause,” “nucleic acid testing,” “Wuhan pneumonia,” “human-to-human transmission,” and the following in English: “COVID-19,”<br>“COVID,” “COVID19,” “SARS,” “SARS-2,” “SARS-CoV-2,” “nCoV,” “2019-nCoV.” The time interval was from 00:00:00 29<br>April 2021 to 04:00:00 18 May 2021. It captures the activity of 1,052,896 users, who have 3,841,030 following relationship edges<br>among them.</p> <p>To protect the privacy of social media users, only the essential variables necessary for constructing post, comment, repost, and follower relationships were retained, while other variables that could potentially compromise user privacy have been removed from the new dataset.</p> <div>*all id's are in the form of Universally Unique Identifier (UUID)</div> <div> </div> <div>File name --- description:</div> <div> </div> <div>* tweet_spider_by_tweet_id_uuid.json --- json list data file of the orignial post:</div> <div> </div> <div>weibo_id_uuid: weibo post id,<br>user_uuid: post author user id,<br>creat_at_h: Weibo post publish time rounded down to hourly level</div> <div> <p>* comment.json --- Comment metadata of a Weibo post:</p> <p>user_uuid: commenter user id <br>ori_weibo_id_uuid: original post weibo id <br>creat_at_h: Weibo post publish time rounded down to hourly level </p> <p>* repost.jsonl --- Repost metadata of a Weibo post. Repost is a special type of Weibo post, with repost relation between original post and repost. The data structure of repost is the same as post.</p> </div> <div> <p>user_uuid: repost author user id<br>ori_weibo_id_uuid: original post weibo id<br>creat_at_h: Weibo post publish time rounded down to hourly level</p> </div> <div> </div> <div>* follower.jsonl --- json list data file of the follower profiles of users. This dataset describes the following relationship, in which fan refers to the user who follower others, and follower refers to the targer user who receive other user's following: <p>fan -> follower<br>fan_id_uuid: fan user's id<br>follower_id_uuid: follower user's id</p> </div> <p> </p> <p> </p>
Weibo and PHEME
Open the record for dataset details and reuse information.
Weibo_hashtags_thematic_analysis_COVID-19
<p>This dataset is affiliated to the journal article: Xi, W., Xu, W., Zhang, X., Ayalon, L. A Thematic Analysis of Weibo Topics (Chinese Twitter Hashtags) Regarding Older Adults During the COVID-19 Outbreak, <em>The Journals of Gerontology: Series B</em>, Volume 76, Issue 7, September 2021, Pages e306–e312, <a href="https://doi.org/10.1093/geronb/gbaa148">https://doi.org/10.1093/geronb/gbaa148</a></p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.