Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

19

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

19 results for “treebank”

Learn how ShareScore rates datasets ↗
zenodo44/100

Late Latin Charter Treebank 1 (LLCT1), version 1.2

<p>Version 1.2 of the Late Latin Charter Treebank 1 (LLCT1). Contains a number of minor corrections, replaces the version 1.0 published at Zenodo in 2018. Early Medieval Latin documentary texts from Italy between AD 714-869 with morphological and syntactic annotation. Latin Dependency Treebank (LDT) compatible linguistic annotation, Prague style treebank format (PML). For a detailed description of the Late Latin Charter Treebanks, see the pre-print of the paper &#39;Late Latin Charter Treebank: contents and annotation&#39;, to be published in Corpora, 16:2 (2021), at the <a href="https://researchportal.helsinki.fi/fi/publications/late-latin-charter-treebank-contents-and-annotation">institutional repository of the University of Helsinki</a>. See also Korkiakangas, T. and Lassila, M. (2013), <a href="https://www.academia.edu/5491302/Korkiakangas_Timo_and_Lassila_Matti_Abbreviations_fragmentary_words_formulaic_language_treebanking_mediaeval_charter_material_"><em>Abbreviations, fragmentary words, formulaic language: treebanking medieval charter material</em></a>, in Mambrini, F., Passarotti, M. and Sporleder, C., <em>Proceedings of the third workshop on annotation of corpora for research in the humanities</em>, pp. 61&ndash;72, and Korkiakangas, T. and Passarotti, M. (2011), <a href="https://pdfs.semanticscholar.org/6825/a8ad70fe6e2a77540d9ff2774b8f34804fd0.pdf?_ga=2.234021022.1247020789.1580545874-419947753.1580545874"><em>Challenges in Annotating Medieval Latin Charters</em></a>, in &laquo;Journal of Language Technology and Computational Linguistics&raquo;, 26, pp. 103&ndash;114.</p>

opencc-by-4.0Jan 2020View details →
zenodo44/100

Treebank of Artemidorus Onir. book V

<p>Collection of Ancient Greek annotated trees&nbsp;of Artemidorus&#39; Oneirocritica Book 5.&nbsp; Part of the Open Projects in Digital Classics at the College of Letters and Sciences of the State University of S&atilde;o Paulo in Araraquara, S&atilde;o Paulo, Brazil.<br> <br> The trees were annotated manually on Perseids Platform using the Arethusa tool. The treebank tagset and guidelines used were those from&nbsp;<a href="https://github.com/PerseusDL/treebank_data/blob/master/AGDT2/guidelines/Greek_guidelines.md">The Ancient Greek Dependency Treebank&nbsp;</a>with the morphological and syntactic layer, which was based on Bamman&#39;s and Crane&#39;s 2008&nbsp;<em>Guidelines for the Syntactic Annotation of the Ancient Greek Dependency Treebank (1.1).&nbsp;</em>We translated it into Portuguese with a few additions from some specifications provided by a forum maintained in 2013 by Alpheios.net,&nbsp;<a href="https://web.archive.org/web/20160401072609/http://treebank.alpheios.net/book">which is no longer online</a>. The trees are visible&nbsp;in Perseids Collection as UNESP-trees at&nbsp;<a href="https://perseids-publications.github.io/unesp-trees">https://perseids-publications.github.io/unesp-trees</a>.</p>

opencc-by-4.0Feb 2022View details →
zenodo40/100

Late Latin Charter Treebank 2 (LLCT2), version 1.2

<p>Version 1.2 of the Late Latin Charter Treebank 2 (LLCT2). Contains a number of minor corrections and&nbsp;replaces the version 1.0 published at Zenodo in 2019. Early Medieval Latin documentary texts from Italy between AD 774-897 with morphological and syntactic annotation. Latin Dependency Treebank (LDT) compatible linguistic annotation, CoNLL treebank format. Note that LLCT2 is also available open-access in the Universal Dependencies format at the <a href="https://github.com/UniversalDependencies/UD_Latin-LLCT">website</a>&nbsp;of the Universal Dependencies consortium. For a detailed description of the Late Latin Charter Treebanks, see the pre-print of the paper &#39;Late Latin Charter Treebank: contents and annotation&#39;, to be published in Corpora, 16:2 (2021), at the <a href="https://researchportal.helsinki.fi/fi/publications/late-latin-charter-treebank-contents-and-annotation">institutional repository of the University of Helsinki</a>. See also Korkiakangas, T. and Lassila, M. (2013), <a href="https://www.academia.edu/5491302/Korkiakangas_Timo_and_Lassila_Matti_Abbreviations_fragmentary_words_formulaic_language_treebanking_mediaeval_charter_material_"><em>Abbreviations, fragmentary words, formulaic language: treebanking medieval charter material</em></a>, in Mambrini, F., Passarotti, M. and Sporleder, C., <em>Proceedings of the third workshop on annotation of corpora for research in the humanities</em>, pp. 61&ndash;72, and Korkiakangas, T. and Passarotti, M. (2011), <a href="https://pdfs.semanticscholar.org/6825/a8ad70fe6e2a77540d9ff2774b8f34804fd0.pdf?_ga=2.234021022.1247020789.1580545874-419947753.1580545874"><em>Challenges in Annotating Medieval Latin Charters</em></a>, in &laquo;Journal of Language Technology and Computational Linguistics&raquo;, 26, pp. 103&ndash;114.</p>

opencc-by-4.0Jan 2020View details →
zenodo40/100

TuDeT: Tupían Dependency Treebank

TuDeT: Tupían Dependency Treebank

opencc-by-sa-4.0Mar 2021View details →
zenodo40/100

Pedalion Ancient Greek Dependency Treebank - Euripides: Medea

<p>xml treebank</p> <p>Annotated by Toon Van Hal, with student contributions by Mathieu Cuijpers; Sanderijn Gijbels; Yoran Joosten; Yordi Lenaerts; Eva Uffing; Chiara Van der Hasselt; Lisa Vanhee and Jolien Volders (KU Leuven Bachelor 3, 2018-2019). Based on a preparsed text by Alek Keersmaekers. Controlled by Toon Van Hal, Sanderijn Gijbels and Yoran Joosten.</p>

opencc-by-nc-sa-4.0Dec 2018View details →
zenodo40/100

The English Headline Treebank corpus

<p>This repository contains the evaluation sets used in</p> <pre><code>A Benton, T Shi, O İrsoy, and I Malioutov."Weakly Supervised Headline Dependency Parsing". Findings of EMNLP. 2022. </code></pre> <p>This dataset contains parse annotations for English news headlines and a script to produce conllu files joined with original headline text.</p> <p>Parse annotations are joined to the corresponding text by running:</p> <pre> LDC_NYT_DIR=&quot;/PATH/TO/UNTARRED/LDC2008T19/&quot; # path to untarred LDC2008T19 python build_eht.py --nyt_dir ${LDC_NYT_DIR} --num_proc 4</pre> <p>This will download the Google sentence compression (GSC) dataset, and build conllu files for GSC examples. If you have the New York Times Annotated Corpus (LDC2008T19) untarred locally, this will also join annotations to the NYT examples (location passed via&nbsp;<code>--nyt_dir</code>).</p> <p>Increase the argument to&nbsp;<code>--num_procs</code>&nbsp;to process more shards from the NYT corpus in parallel and reduce build time. The above was tested with python 3.9.7.</p> <ul> <li>The EHT evaluation sets, with gold-annotated POS tags and dependency relations, are built as&nbsp;<code>EHT/gsc.test.conllu</code>&nbsp;and&nbsp;<code>EHT/nyt.test.conllu</code></li> <li>Silver, projected, trees which we used to train and validate out models are built under&nbsp;<code>GSC_projected</code>. These are not gold parse trees (projected predictions from the article lead sentence), and are shared purely for reproducibility sake.</li> </ul>

opencc-by-4.0Nov 2022View details →
zenodo40/100

Binary Stanford Sentiment Treebank 2 (SST-2)

<p>Binary Stanford Sentiment Treebank (SST2)&nbsp;is&nbsp; a binary version of SST and Movie Review dataset&nbsp;(the neutral class was removed), that is, the data was classified only into positive and negative classes.</p> <p>The files:<br> texts.txt: Document set (text). One per line.<br> score.txt: Document class whose index is associated with texts.txt<br> split_&lt;k&gt;.pkl:&nbsp;&nbsp;pandas DataFrame with k-cross validation partition</p>

opencc-by-4.0Aug 2021View details →
zenodo40/100

Stanford Sentiment Treebank (SST)

<p>Stanford Sentiment Treebank (SST) is an extension of Movie Review dataset&nbsp;with fine-grained labels ranging between very positive and very negative. The authors&nbsp;extended the MR by adding a more curated human annotation into 5&nbsp;classes.<br> &nbsp;</p> <p>The files:<br> texts.txt: Document set (text). One per line.<br> score.txt: Document class whose index is associated with texts.txt<br> split_&lt;k&gt;.pkl:&nbsp;&nbsp;pandas DataFrame with k-cross validation partition</p>

opencc-by-4.0Aug 2021View details →
zenodo36/100

Berkeley neural parser model for Korean: Sejong treebank

<p>Berkeley Neural Parser model for Korean using the Sejong treebank</p> <p>&nbsp;</p> <p>Berkeley Neural Parser&nbsp;https://github.com/nikitakit/self-attentive-parser&nbsp;</p> <p>COLLINS-SJTREE.prm for evalb is available at&nbsp;http://doi.org/10.5281/zenodo.1004604&nbsp;</p> <p>&nbsp;</p>

opencc-by-4.0Aug 2020View details →
zenodo36/100

Kobalt_RST (RST German Learner Treebank): Die Annotation von rhetorischen Strukturen im Kobalt-DaF-Korpus

<p>Das <a href="https://www.linguistik.hu-berlin.de/de/institut/professuren/korpuslinguistik/forschung/kobalt-daf">Kobalt-DaF-Korpus</a> ist ein systematisch erhobenes und tief annotiertes&nbsp;Deutschlernerkorpus, welches 80 deutschsprachige argumentative Texte von deutschen L1-Sprecher:innen und Deutschlerner:innen&nbsp;unterschiedlicher&nbsp;L1 enth&auml;lt.&nbsp;Dieses Repositorium stellt&nbsp;eine zus&auml;tzliche Annotation des Kobalt-DaF-Korpus&nbsp;bzgl.&nbsp;rhetorischer Strukturen frei zur Verf&uuml;gung.&nbsp;Folgende Informationen sind hier zu finden: (1) Die Darstellung des Annotationsprozesses (Annotationsframework, -richtlinie, und -verfahren). (2) Die annotierten rs3-Dateien.</p> <p>*Versionshinweise:&nbsp;Bislang sind ausschlie&szlig;lich die Texte der chinesischen Deutschlerner:innen und der deutschen L1-Sprecher:innen (insgesamt 40 Texte) verf&uuml;gbar.&nbsp;Die Annotation der &uuml;brigen Texte folgt demn&auml;chst.&nbsp;</p> <p>*Die Annotationsarbeit wurde gef&ouml;rdert durch das Chinese Scholarship Council und die Deutsche Forschungsgemeinschaft (DFG) &ndash; SFB 1412, 416591334.</p>

opencc-by-4.0Nov 2021View details →
zenodo36/100

Trance parser model for Korean: Sejong treebank

<p>Trance parsing model + Embedding vector.</p> <p>See https://github.com/tarowatanabe/trance for the parser and its usage. We also provide parsing and learning scripts for the Trance parser that we used for the paper;&nbsp;</p> <p>1/ parsing model:&nbsp;ptb_train.txt.model-d100.tar.gz &nbsp; &nbsp; &nbsp; &nbsp;</p> <p>2/ embedding vector:&nbsp;embedding-d100.vec.gz</p> <p>3/ trance parser parsing script:&nbsp;trance-parsing.sh</p> <p>4/ trance parser (batch) learning&nbsp;script:&nbsp;trance-training-batch.sh</p> <p>5/ test.txt (gold file) and test.txt.leaf is for the parser input.&nbsp;</p> <p>&nbsp;</p>

opencc-by-4.0Oct 2017View details →
zenodo36/100

Book One of The Iliad, Persian and Kurdish Translation with Word-level Alignment to the Treebank, Didakta Annotations and Glossary

<p>This dataset includes Persian and Kurdish translation of book one of the Iliad, aligned at word-level with the treebanks. The Persian translation is by Farnoosh Shamsian and the Kurdish translation by Farshid Rahimi. The dataset also includes glossaries in both languages.&nbsp;</p> <p>The treebank data combines both the UD treebank and the Perseus treebank in one spreadsheet. The Perseus treebank is available here:&nbsp;<a href="http://perseusdl.github.io/treebank_data/">http://perseusdl.github.io/treebank_data/</a></p> <p>The UD version is converted by Francesco Mambrini. For more information, see:&nbsp;<a href="https://github.com/francescomambrini/katholou/tree/main/ud_treebanks/agdt/data">https://github.com/francescomambrini/katholou/tree/main/ud_treebanks/agdt/data</a></p> <p>Both translations are aligned word by word to the treebank according to guidelines. The guideline for the Persian alignment is available here:</p> <p>Farnoosh Shamsian. (2023). Alignment Guidelines for Classical Greek-Persian. Zenodo. <a href="https://doi.org/10.5281/zenodo.8039931">https://doi.org/10.5281/zenodo.8039931</a></p> <p>The spreadsheet also contains Didakta annotations. For moe information about Didakta, see:&nbsp;</p> <p>Farnoosh Shamsian. (2023). Didakta Grammar for Annotation (English). Zenodo. <a href="http://https://doi.org/10.5281/zenodo.8318137">https://doi.org/10.5281/zenodo.8318137</a></p> <p>&nbsp;</p>

opencc-by-4.0Sep 2023View details →
zenodo32/100

PapyGreek Treebanks: First prerelease

<p>This is the first prerelease of PapyGreek annotations. The full release is located at&nbsp;<a href="https://doi.org/10.5281/zenodo.5062996">https://doi.org/10.5281/zenodo.5062996</a>.</p>

opencc-by-4.0Dec 2020View details →
zenodo32/100

MaltParser model for Korean: Sejong treebank

<p>MaltParser model for Korean:  Sejong treebank</p> <p>Jungyeul Park, Jeen-Pyo Hong, and Jeong-Won Cha (2016). Korean Language Resources for Everyone. In Proceedings of the 30th Pacific Asia Conference on Language, Information and Computation (PACLIC 30). Seoul, Korea. [pdf]</p> <p>@inproceedings{park-hong-cha:2016:PACLIC, <br> address = {Seoul, Korea}, <br> author = {Park, Jungyeul and Hong, Jeen-Pyo and Cha, Jeong-Won}, <br> booktitle = {Proceedings of the 30th Pacific Asia Conference on Language, Information and Computation (PACLIC 30)}, <br> pages = {49--58}, <br> title = {{Korean Language Resources for Everyone}}, <br> year = {2016} <br> }</p> <p>It requires Espresso's POS tagging results for input. Espresso is available at https://zenodo.org/record/884606 </p>

opencc-by-sa-4.0Sep 2017View details →
zenodo32/100

Burmese (Myanmar) Treebank of Asian Language Treebank Project

<p>* Introduction</p> <p>This is the Myanmar ALT of the Asian Language Treebank (ALT) Corpus. Please refer to<br> http://www2.nict.go.jp/astrec-att/member/mutiyama/ALT/index.html<br> for an introduction of the ALT project.</p> <p>The process of building the Myanmar ALT began with sampling about 20,000 sentences from English Wikinews, and then these sentences were translated into Myanmar language.<br> &nbsp;&nbsp; &nbsp;<br> The English Wikinews<br> https://en.wikinews.org/wiki/Main_Page<br> is available under the terms of the Creative Commons Attribution 2.5 License.<br> https://creativecommons.org/licenses/by/2.5/</p> <p>Myanmar ALT has been developed by NICT and UCSY. The license of Myanmar ALT is</p> <p>Creative Commons Attribution-NonCommercial-ShareAlike 4.0 International (CC BY-NC-SA 4.0) License<br> https://creativecommons.org/licenses/by-nc-sa/4.0/</p> <p><br> * Contents</p> <p>- data : Myanmar ALT treebank</p> <p>&nbsp;</p>

opencc-by-nc-sa-4.0Sep 2019View details →
zenodo28/100

Khmer Treebank of Asian Language Treebank Project

<p>* Introduction</p> <p>This is the Khmer ALT of the Asian Language Treebank (ALT) Corpus. Please refer to<br> http://www2.nict.go.jp/astrec-att/member/mutiyama/ALT/index.html<br> for an introduction of the ALT project.</p> <p>The process of building the Khmer ALT began with sampling about 20,000 sentences from English Wikinews, and then these sentences were translated into Khmer language.<br> &nbsp;&nbsp; &nbsp;<br> The English Wikinews<br> https://en.wikinews.org/wiki/Main_Page<br> is available under the terms of the Creative Commons Attribution 2.5 License.<br> https://creativecommons.org/licenses/by/2.5/</p> <p>* License</p> <p>Khmer ALT has been developed by NICT and CADT (a.k.a. NIPTICT). The license of Khmer ALT is</p> <p>Academic Research Non-Commercial Limited CC-BY-NC-SA Reference-Type License Agreement Terms and Conditions</p> <p>The use of this material (copyrighted material, data, etc.) is permitted under the conditions similar to the Creative Commons &quot;Attribution-NonCommercial-ShareAlike&quot; License. However, the purpose of use shall not only be &quot;NonCommercial&quot; but must be &ldquo;for Academic Research and Non-Commercial&rdquo;.<br> The detailed legal provisions are listed at the end of the document, but the outline is as follows.</p> <p>You are free to:<br> Share &mdash; copy and redistribute the material in any medium or format<br> Adapt &mdash; remix, transform, and build upon the material<br> The licensor cannot revoke these freedoms as long as you follow the license terms.</p> <p>Under the following terms:<br> Attribution &mdash; You must give appropriate credit, provide a link to the license, and indicate if changes were made. You may do so in any reasonable manner, but not in any way that suggests the licensor endorses you or your use.<br> Academic Research&amp;Non-Commercial &mdash; You may use the material only for academic research purposes and may not use it for commercial purposes.<br> ShareAlike &mdash; If you remix, transform, or build upon the material, you must distribute your contributions under the same license as the original.</p> <p>In this Terms and Conditions, academic research purposes are not associated with the interpretation of the Creative Commons License, and shall refer to &quot;the purpose of intellectual creative activities conducted by persons belonging to institutions of higher education and research institutes, etc. in order to develop a better and affluent society through exploration of the truth about nature, humans, society, etc., discovery of new principles and laws, and application of these in a wide range of fields from natural sciences to human and social sciences.&quot; Non-commercial means: not for the purpose of any contribution to a for-profit or commercial business. The outcome of research and development activities initially undertaken without the purpose of contributing in any way to a for-profit or commercial business shall not be permitted to be used later to contribute to one.</p> <p>The legal provisions of this license are: the terms of Creative Commons Attribution-NonCommercial-ShareAlike 4.0 International (https://creativecommons.org/licenses/by-nc-sa/4.0/legalcode) plus the following rider clauses. On that note, the Terms and Conditions of this license are not equal to that of the Creative Commons license&mdash;they were inspired by them and share most of the terms and conditions with them, but shall be considered as a different license.</p> <p>Rider Clauses<br> 1. &quot;Creative Commons Attribution-NonCommercial-ShareAlike 4.0 International Public License&quot; in the title and preamble shall read &quot;Academic Research Non-Commercial Limited CC-BY-NC-SA Compliant License Agreeement Terms and Conditions&quot;.<br> 2. Article 1 c and g shall be deleted.<br> 3. Article 1 k shall be amended as follows:<br> k. &quot;Academic Research Purposes&quot; shall refer to the purpose of intellectual creative activities conducted by persons belonging to institutions of higher education and research institutes, etc. in order to develop a better and affluent society through exploration of the truth about nature, humans, society, etc., discovery of new principles and laws, and application of these in a wide range of fields from natural sciences to human and social sciences, and does not consider contribution to for-profit or commercial businesses as its main or secondary purpose. Activities initially undertaken without the purpose of contributing to a for-profit or commercial business, shall no longer be considered as &quot;academic research purposes&quot; as soon as they decide to contribute to one.<br> 4. Article 3 b.1. shall be replaced by the following:<br> 1. The Adapter&#39;s License that You apply must be this License.<br> 5. In all provisions, &quot;NonCommercial&quot; shall read &quot;Academic Research Purposes&quot; and &quot;this Public License&quot; shall read &quot;this License&quot;.</p> <p>Creative Commons Attribution-NonCommercial-ShareAlike 4.0 International Public License<br> &lt;https://creativecommons.org/licenses/by-nc-sa/4.0/legalcode&gt;</p> <p>* Usage</p> <p>1. Compile the source codes.<br> 2. Run with no option to see the instruction.<br> 3. Follow the instruction to generate the treebank.</p>

opencc-by-4.0Jul 2023View details →
zenodo24/100

Tokenized and POS-Tagged Khmer Data of the Asian Language Treebank Project

<p>* Introduction</p> <p>This is the Khmer ALT of the Asian Language Treebank (ALT) Corpus. English texts sampled from English Wikinews were available under a Creative Commons Attribution 2.5 License.</p> <p>Please refer to<br> http://www2.nict.go.jp/astrec-att/member/mutiyama/ALT/index.html<br> for an introduction of the ALT project.</p> <p>Khmer ALT has been developed by NICT and NIPTICT. The license of Khmer ALT is</p> <p>Creative Commons Attribution-NonCommercial-ShareAlike 4.0 International (CC BY-NC-SA 4.0) License<br> https://creativecommons.org/licenses/by-nc-sa/4.0/</p> <p><br> * Contents</p> <p>- data_km.km-[tok|tag].nova : tokenized/POS-tagged Khmer sentences by the nova annotation system<br> &nbsp;&nbsp; &nbsp;# based on the following two guildelines<br> &nbsp;&nbsp; &nbsp;# http://www2.nict.go.jp/astrec-att/member/mutiyama/ALT/Khmer-annotation-guideline.pdf<br> &nbsp;&nbsp; &nbsp;# http://www2.nict.go.jp/astrec-att/member/mutiyama/ALT/Khmer-annotation-guideline-supplementary.pdf</p> <p><br> * Disclaimer</p> <p>[1] The content of the selected English Wikinews articles have been translated for this corpus. English texts sampled from English Wikinews were available under a Creative Commons Attribution 2.5 License. Users of the corpus are requested to take careful consideration when encountering any instances of defamation, discriminatory terms, or personal information that might be found within the corpus. Users of the corpus are advised to read Terms of Use in https://en.wikinews.org/wiki/Main_Page carefully to ensure proper usage.</p> <p>[2] NICT bears no responsibility for the contents of the corpus and the lexicon and assumes no liability for any direct or indirect damage or loss whatsoever that may be incurred as a result of using the corpus or the lexicon.</p> <p>[3] If any copyright infringement or other problems are found in the corpus or the lexicon, please contact us at alt-info[at]khn[dot]nict[dot]go[dot]jp. We will review the issue and undertake appropriate measures when needed.</p>

opencc-by-4.0Jul 2020View details →
zenodo24/100

Berkeley parser model for Korean: Sejong treebank

<p><strong>Berkeley parser model for Korean:  Sejong treebank</strong></p> <p>Jungyeul Park, Jeen-Pyo Hong, and Jeong-Won Cha (2016). Korean Language Resources for Everyone. In Proceedings of the 30th Pacific Asia Conference on Language, Information and Computation (PACLIC 30). Seoul, Korea. [pdf]</p> <p>@inproceedings{park-hong-cha:2016:PACLIC, <br> address = {Seoul, Korea}, <br> author = {Park, Jungyeul and Hong, Jeen-Pyo and Cha, Jeong-Won}, <br> booktitle = {Proceedings of the 30th Pacific Asia Conference on Language, Information and Computation (PACLIC 30)}, <br> pages = {49--58}, <br> title = {{Korean Language Resources for Everyone}}, <br> year = {2016} <br> }</p> <p>It requires Espresso's POS tagging results for input. Espresso is available at https://zenodo.org/record/884606 </p>

opencc-by-sa-4.0Sep 2017View details →
zenodo24/100

Berkeley parser model for Korean: Sejong treebank

<p>Berkeley parser model for Korean: Sejong treebank (2020 results) requires scripts from&nbsp;http://doi.org/10.5281/zenodo.3995082</p> <p>&nbsp;</p>

opencc-by-4.0Sep 2017View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record