Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

2

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

2 results for “Chinese local gazetter”

Learn how ShareScore rates datasets ↗
zenodo40/100

Mulberry Disasters in Chinese Local Gazetteers

<p>This dataset contains 404 mulberry disasters found in a digital collection of 4,000 Chinese local gazetteers (published in Erudition&#39;s Zhongguo Fangzhi Ku, or the Database of Chinese Local Gazetteers) that was curated during 2018 and 2019 by scholars at the Max Planck Institute for the History of Science to support their joint research paper entitled: &ldquo;What Is Local Knowledge: Digital Humanities and Yuan Dynasty Disasters in Imperial China&rsquo;s Local Gazetteers.&rdquo; The paper appeared in the <em>Journal of Chinese History</em> in its 2020 spring issue. The authors use this dataset to ask what results the emerging methodology of analyzing data drawn from historical sources can produce for historical research, given that such data-driven analysis is inevitably quantitative and that many historians believe this would contradict the core value of studies in the humanities.</p> <p>This dataset is accompanied by a data paper to be published by Brill&#39;s Digital Concordances Platform. In the data&nbsp;paper, the authors give information about the full curation cycle of this dataset, including the research goal that motivated the curation, the sources that were used, and the curation/cleansing/preparation process. The authors also explain this dataset in detail, including the kind of information it contains and its overall temporal and geospatial distributions. In conclusion, they discuss the potential usages of the dataset. With this paper, we also hope to offer an exemplary data curation workflow for collecting data from historical sources that includes not only data curation (whether manual or semiautomatic), but also the practical steps of data cleansing and normalization that reflect the various historical considerations in the process.</p>

opencc-by-4.0May 2022View details →
zenodo40/100

N-gram dataset of Chinese local gazetteers (中國地方誌)

<p>This dataset contains the N-grams (1-3) collected from 11083 Chinese local gazetteers&nbsp; (中國地方誌).</p> <p>The dataset comprises of the following resources:</p> <ul> <li><strong>local_gazetteer_1.7z</strong> Unigram dataset in tab separated format (one file per book, each row contains the N-gram and its count)</li> <li><strong>local_gazetteer</strong><strong>_2.7z</strong> Bigram dataset in tab separated format&nbsp;(one file per book, each row contains the N-gram and its count)</li> <li><strong>local_gazetteer</strong><strong>_3.7z</strong> Trigram dataset in tab separated format&nbsp;(one file per book, each row contains the N-gram and its count)</li> <li><strong>local_gazetteer</strong><strong>_metadata.xlsx</strong> Metadata of each book</li> </ul> <p>&nbsp;</p> <p>Dieses Datenset enth&auml;lt die in 11083 chinesischen Lokalmonographien (中國地方誌) enthaltenen&nbsp;N-Gramme (1-3).&nbsp;&nbsp;</p> <p>Das Datenset besteht aus den folgenden Dateien:</p> <ul> <li><strong>local_gazetteer</strong><strong>_1.7z</strong><em> </em>Monogramm-Datenset im .txt Dateiformat&nbsp;mit Tabstopp als&nbsp;Trennzeichen (jede Datei enth&auml;lt ein Buch, jede Zeile ein N-Gramm mit der Anzahl der Vorkommnisse im Text)</li> <li><strong>local_gazetteer</strong><strong>_2.7z</strong><em>&nbsp;</em>Bigramm-Datenset im .txt Dateiformat&nbsp;mit Tabstopp als&nbsp;Trennzeichen (jede Datei enth&auml;lt ein Buch, jede Zeile ein N-Gramm mit der Anzahl der Vorkommnisse im Text)</li> <li><strong>local_gazetteer</strong><strong>_3.7z</strong><em> </em>Trigramm-Datenset im .txt Dateiformat&nbsp;mit Tabstopp als&nbsp;Trennzeichen (jede Datei enth&auml;lt ein Buch, jede Zeile ein N-Gramm mit der Anzahl der Vorkommnisse im Text)</li> <li><strong>local_gazetteer</strong><em><strong>_</strong></em><strong>metata.xlsx</strong> Metadaten der enthaltenen B&uuml;cher</li> </ul> <p>&nbsp;</p> <p>11083 中國地方誌n元語法統計資料 (N-gram Dataset)</p> <p>以下是檔案簡說:</p> <ul> <li><strong>local_gazetteer</strong><strong>_1.7z</strong><em>&nbsp;</em>中國地方誌一元分詞(Unigram)的統計資料 (每本書一個檔案, 以tab作欄區分, 每一行紀錄該N-gram在書中出現的次數)</li> <li><strong>local_gazetteer</strong><strong>_2.7z</strong><em> </em>中國地方誌二元分詞(Bigram)的統計資料&nbsp;(每本書一個檔案, 以tab作欄區分, 每一行紀錄該N-gram在書中出現的次數)</li> <li><strong>local_gazetteer</strong><strong>_3.7z</strong><em> </em>中國地方誌三元分詞(Trigram)的統計資料 (每本書一個檔案, 以tab作欄區分, 每一行紀錄該N-gram在書中出現的次數)</li> <li><strong>local_gazetteer</strong><em><strong>_</strong></em><strong>metadata.xlsx</strong> 紀錄每本書的基本Metadata</li> </ul>

opencc-by-4.0Mar 2019View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record