Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

67

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

67 results for “Captioning”

Learn how ShareScore rates datasets ↗
zenodo48/100

Videos, audio transcriptions and closed captions for the workshop: Introduction to Wikidata for Maastricht University, Theory and Pratice - 15 October 2024

<h1><strong>Videos, audio transcriptions and closed captions for the workshop:&nbsp; Introduction to Wikidata for Maastricht University, Theory and Pratice - 15 October 2024</strong></h1> <h3><a href="https://theplant.maastrichtuniversity.nl/event/navigating-the-world-of-wikidata-for-research-science-and-cultural-heritage-2"><em>Navigating the World of Wikidata for Research, Science and Cultural Heritage</em></a></h3> <h3><em><a href="https://www.wikidata.org/wiki/Wikidata:Twelfth_Birthday/Workshop_in_Maastricht" target="_blank" rel="noopener">Wikidata's Twelfth Birthday: Workshop in Maastricht&nbsp;</a></em></h3> <p>Wikidata is a free, collaborative, multilingual database, collecting structured open data for anyone in the world to use. It also plays a crucial role in supporting Wikimedia projects, such as Wikipedia and Wikimedia Commons. Over the last 12 years it has strongly increased in popularity among the scientific and cultural heritage communities.</p> <p>In this 2,5 hours workshop you will learn the basics of working with Wikidata, both in theory and practice. You will learn</p> <ol> <li>The basics of Wikidata: A first look at what Wikidata is and how it works, both technically and socially (Wikidata community)</li> <li>How Wikidata can be relevant for research, science and cultural heritage (GLAM), and</li> <li>First steps in contributing to Wikidata yourself, with a focus on the topic of UM professors from past and present.</li> </ol> <p>As part of the <a title="Wikidata:Twelfth Birthday" href="https://www.wikidata.org/wiki/Wikidata:Twelfth_Birthday">Wikidata 12th Birthday celebrations</a> this workshop is open to academics, researchers, students, and professionals interested in working with Wikidata in the intersection of open data, research, and science. Whether you are new to Wikidata or looking to deepen your understanding, this session will provide valuable insights for improving your work.</p> <h2><strong>Workshop outline</strong></h2> <h3><strong>Part 1:&nbsp; Theory, Wikidata basics (45-60 minutes) </strong></h3> <ul> <li><strong><a href="https://zenodo.org/records/13984149/files/Wikidata%20Workshop%20-%20Theoretical%20part%20-%20Maastricht%20University%20-%2015%20October%202024.webm" target="_blank" rel="noopener">Video</a> (.webm) including&nbsp;<a href="https://zenodo.org/records/13984149/files/WikidataWorkshop_MaastrichtUniversity_15October2024_TheoreticalPart.txt?download=1" target="_blank" rel="noopener">audio transcription </a>(.txt) and <a href="https://zenodo.org/records/13984149/files/Wikidata%20Workshop%20-%20Theoretical%20part%20-%20Maastricht%20University%20-%2015%20October%202024.webm.en.srt?download=1" target="_blank" rel="noopener">closed captions</a> (.srt) are available below</strong></li> </ul> <p><strong>Additional materials</strong></p> <ul> <li><strong>Slides in&nbsp;<a href="https://zenodo.org/records/13837957/files/WikidataWorkshop_MaastrichtUniversity_15October2024_TheoreticalPart.pptx?download=1" rel="nofollow">PowerPoint</a> or <a href="https://zenodo.org/records/13837957/files/Wikidata%20Workshop%20-%20Theoretical%20part%20-%20Maastricht%20University%20-%2015%20October%202024.pdf?download=1" rel="nofollow">PDF</a> are available from <a href="https://zenodo.org/records/13837957" target="_blank" rel="noopener">https://zenodo.org/records/13837957</a></strong></li> </ul> <p><em>1) Wikidata basics</em></p> <ul> <li>What is Wikidata?</li> <li>What are the principles of Wikidata?</li> <li>How are things described in Wikidata?</li> <li>Who builds Wikidata? - The Wikidata community</li> </ul> <p><em>2) Wikidata for research, science and cultural heritage</em></p> <ul> <li>To what extent is Wikidata used throughout science, research and GLAM?</li> <li>Six anecd<em>a</em>tic cases <ol> <li>Scientometrics - Scholia</li> <li>Life and biomedical sciences</li> <li>Astronomy</li> <li>Language technology / AI / LLMs</li> <li>GLAM &ndash; KB collection highlights</li> <li>Representation of (female) scientists</li> </ol> </li> </ul> <h3><strong>Break (15 minutes)</strong></h3> <h3><strong>Part 2: Practice, contributing to Wikidata (75-90 minutes)&nbsp;</strong></h3> <ul> <li><strong><a href="https://zenodo.org/records/13984149/files/Wikidata%20Workshop%20-%20Practical%20part,%20UM%20professors%20-%20Maastricht%20University%20-%2015%20October%202024.webm?download=1" target="_blank" rel="noopener">Video</a> (.webm) including <a href="https://zenodo.org/records/13984149/files/WikidataWorkshop_MaastrichtUniversity_15October2024_PracticalPart_UMprofessors.txt?download=1" target="_blank" rel="noopener">audio transcription </a>(.txt) and <a href="https://zenodo.org/records/13984149/files/Wikidata%20Workshop%20-%20Practical%20part,%20UM%20professors%20-%20Maastricht%20University%20-%2015%20October%202024.webm.en.srt?download=1" target="_blank" rel="noopener">closed captions</a> (.srt) are available below</strong></li> </ul> <p><strong>Additional materials</strong></p> <ul> <li><strong>Slides in&nbsp;<a href="https://zenodo.org/records/13837957/files/WikidataWorkshop_MaastrichtUniversity_15October2024_PracticalPart_UMprofessors.pptx?download=1" rel="nofollow">PowerPoint</a> or <a href="https://zenodo.org/records/13837957/files/Wikidata%20Workshop%20-%20Practical%20part,%20UM%20professors%20-%20Maastricht%20University%20-%2015%20October%202024.pdf?download=1" rel="nofollow">PDF</a> are available from <a href="https://zenodo.org/records/13837957" target="_blank" rel="noopener">https://zenodo.org/records/13837957</a></strong></li> <li><strong>Handout for participants in <a href="https://zenodo.org/records/13837957/files/WikidataWorkshop_MaastrichtUniversity_15October2024_PracticalPart_HandoutForParticipants.docx?download=1" rel="nofollow">Word</a> or <a href="https://zenodo.org/records/13837957/files/WikidataWorkshop_MaastrichtUniversity_15October2024_PracticalPart_HandoutForParticipants.pdf?download=1" rel="nofollow">PDF</a></strong> <strong>are available from <a href="https://zenodo.org/records/13837957" target="_blank" rel="noopener">https://zenodo.org/records/13837957</a></strong></li> </ul> <p><strong>&nbsp; </strong>The goals of this hands-on part are:</p> <ul> <li>Get familiar with basic data editing via the Wikidata interface</li> <li>Understand WD data models and structures related to professors (of Maastricht University)</li> <li>Extend existing <a title="Wikidata:Wiki-wetenschappers/Universiteit Maastricht/hoogleraren" href="https://www.wikidata.org/wiki/Wikidata:Wiki-wetenschappers/Universiteit_Maastricht/hoogleraren">Wikidata items about UM professors</a>, based on information in public sources.</li> <li>If time allows: Create new Wikidata items about UM professors</li> </ul> <p>The visual slides and the textual handout explain the same content, blocks and exercises, albeit in a slightly different order.<strong>&nbsp; &nbsp; </strong></p> <h2><strong>Required preparation</strong></h2> <p>To make optimal use of our time, participants must create a Wikidata account in the weeks before the workshop. See <a href="https://www.wikidata.org/w/index.php?title=Special:CreateAccount" target="_blank" rel="noopener">https://www.wikidata.org/w/index.php?title=Special:CreateAccount</a>.</p> <p>This is important because very fresh accounts may have limited editing rights. Furthermore only 6 Wikidata accounts can be created per day from UM IP addresses, so creating a lot of new accounts during the workshop might overstretch this limit.</p> <h2><strong>Workshop leader</strong></h2> <p>This workshop was given by <a href="https://www.kb.nl/over-ons/experts/olaf-janssen">Olaf Janssen</a>, the Wikimedia coordinator of the <a href="https://www.kb.nl/over-ons/experts/olaf-janssen">Koninklijke Bibliotheek</a>, the national library of the Netherlands.</p> <p>In this role he stimulates and facilitates collaboration between the collections, knowledge, open data and staff of the KB on the one hand, and the projects of the Wikimedia movement, such as Wikipedia, Wikimedia Commons and Wikidata on the other. He is also active as a volunteer within the community. Feel free to contact Olaf via olaf.janssen(at)<a href="http://kb.nl">kb.nl</a></p> <h2><strong>Materials on Wikimedia Commons</strong></h2> <p>Photos , videos and presentations related to this event can be found on Wikimedia Commons:&nbsp;<a title="c:Category:Wikidata Workshop at Maastricht University, 15 October 2024" href="https://commons.wikimedia.org/wiki/Category:Wikidata_Workshop_at_Maastricht_University,_15_October_2024">Category:Wikidata Workshop at Maastricht University, 15 October 2024</a></p> <h2>Relevant URLs&nbsp;</h2> <ul> <li><a href="https://www.wikidata.org/wiki/Wikidata:Twelfth_Birthday/Workshop_in_Maastricht">https://www.wikidata.org/wiki/Wikidata:Twelfth_Birthday/Workshop_in_Maastricht&nbsp;</a></li> <li><a href="https://theplant.maastrichtuniversity.nl/event/navigating-the-world-of-wikidata-for-research-science-and-cultural-heritage-2" target="_blank" rel="noopener">https://theplant.maastrichtuniversity.nl/event/navigating-the-world-of-wikidata-for-research-science-and-cultural-heritage-2</a>&nbsp; + <a href="https://web.archive.org/web/20240926154021/https://theplant.maastrichtuniversity.nl/event/navigating-the-world-of-wikidata-for-research-science-and-cultural-heritage-2/">archived version</a></li> <li><a href="https://www.linkedin.com/feed/update/urn:li:activity:7244614063102529536/">https://www.linkedin.com/feed/update/urn:li:activity:7244614063102529536/</a></li> <li><a href="https://www.linkedin.com/feed/update/urn:li:activity:7245329234385059843/" target="_blank" rel="noopener">https://www.linkedin.com/feed/update/urn:li:activity:7245329234385059843/</a></li> </ul> <p>Earlier LinkedIn posts (April-May 2024, before rescheduling the worlshop to October)</p> <ul> <li><a href="https://www.linkedin.com/feed/update/urn:li:activity:7188470390858428416/" target="_blank" rel="noopener">https://www.linkedin.com/feed/update/urn:li:activity:7188470390858428416/</a></li> <li><a href="https://www.linkedin.com/feed/update/urn:li:activity:7188180802021543936/" target="_blank" rel="noopener">https://www.linkedin.com/feed/update/urn:li:activity:7188180802021543936/</a></li> <li><a href="https://www.linkedin.com/feed/update/urn:li:activity:7189276985267826691/" target="_blank" rel="noopener">https://www.linkedin.com/feed/update/urn:li:activity:7189276985267826691/</a></li> </ul> <h3>&nbsp;</h3>

opencc-by-4.0Oct 2024View details →
zenodo44/100

Image captioning dataset for human activities

<p>An image captioning dataset including images of humans performing various activities. The included images include the following activities: <code>walking, running, sleeping, swimming, sitting, jumping, riding, climbing, drinking and reading.</code></p>

opencc-by-4.0Jan 2021View details →
zenodo44/100

Swahili Image Captioning Dataset

<p>The SwaFlickr8k dataset is an extension of the well-known Flickr8k dataset, specifically designed for image captioning tasks. It includes a collection of images and corresponding captions written in Swahili. With 8,091 unique images and 40,455 captions, this dataset provides a valuable resource for research and development in the field of image understanding and language processing, particularly in the context of Swahili language.</p>

opencc-by-4.0Jun 2023View details →
zenodo40/100

Audio captioning DCASE 2020 evaluation (testing) split

<p>This is the <strong>evaluation split for Task 6, Automated Audio Captioning, in DCASE 2020 Challenge</strong>.&nbsp;</p> <p>This evaluation split is the Clotho testing split, which is thoroughly described in the corresponding paper:&nbsp;</p> <p><em>K. Drossos, S. Lipping and T. Virtanen, &quot;Clotho: an Audio Captioning Dataset,&quot; IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), Barcelona, Spain, 2020, pp. 736-740, doi: 10.1109/ICASSP40776.2020.9052990.</em></p> <p>available online at: <a href="https://arxiv.org/abs/1910.09387">https://arxiv.org/abs/1910.09387</a> and at: <a href="https://ieeexplore.ieee.org/document/9052990 ">https://ieeexplore.ieee.org/document/9052990&nbsp;</a></p> <p>This evaluation split is meant to be used for the purposes of the Task 6 at the scientific challenge&nbsp;DCASE 2020. This split it is not meant to be used for developing audio captioning methods. For developing audio captioning methods, you should use the development and evaluation splits of Clotho.&nbsp;</p> <p>If you want the development and evaluation splits of Clotho dataset, you can find them also in Zenodo, at: <a href="https://zenodo.org/record/3490684">https://zenodo.org/record/3490684</a></p> <p>--------------------------------------------------------------------------------------------------------</p> <p><strong>== License ==</strong></p> <p>The audio files in the archives:</p> <ul> <li>clotho_audio_test.7z&nbsp;</li> </ul> <p>and the associated meta-data in the CSV file:</p> <ul> <li>clotho_metadata_test.csv</li> </ul> <p>are under the corresponding licences (mostly CreativeCommons with attribution) of Freesound [1] platform, mentioned explicitly in the CSV file&nbsp;for each of the audio files. That is, each audio file in the 7z archive&nbsp;is listed in the CSV file&nbsp;with the meta-data. The meta-data for each file are:&nbsp;</p> <ul> <li>File name</li> <li>Start and ending samples for the excerpt that is used in the Clotho dataset</li> <li>Uploader/user in the Freesound platform (manufacturer)</li> <li>Link to the licence of the file</li> </ul> <p>--------------------------------------------------------------------------------------------------------</p> <p><strong>== References ==</strong><br> [1]&nbsp;Frederic Font, Gerard Roma, and Xavier Serra. 2013. Freesound technical demo. In Proceedings of the 21st ACM international conference on Multimedia (MM &#39;13). ACM, New York, NY, USA, 411-412. DOI: https://doi.org/10.1145/2502081.2502245</p>

openother-atMay 2020View details →
zenodo40/100

IPBES Invasive Alien Species Assessment: Summary for Policymakers. Figures, tables and captions

<p>Figures, tables and their captions from the&nbsp;Summary for Policymakers of the Thematic Assessment Report on Invasive Alien Species and their Control of the Intergovernmental Science-Policy Platform on Biodiversity and Ecosystem Services.</p>

opencc-by-4.0Nov 2023View details →
zenodo40/100

IPBES Invasive Alien Species Assessment: Summary for Policymakers. Figures, tables and captions in Arabic

<p>Arabic translations of the figures, tables and their captions from the Summary for Policymakers of the Thematic Assessment Report on Invasive Alien Species and their Control of the Intergovernmental Science-Policy Platform on Biodiversity and Ecosystem Services.</p>

opencc-by-4.0May 2024View details →
zenodo40/100

IPBES Invasive Alien Species Assessment: Summary for Policymakers. Figures, tables and captions in Chinese

<p>Chinese translations of the figures, tables and their captions from the Summary for Policymakers of the Thematic Assessment Report on Invasive Alien Species and their Control of the Intergovernmental Science-Policy Platform on Biodiversity and Ecosystem Services.</p>

opencc-by-4.0May 2024View details →
zenodo40/100

IPBES Invasive Alien Species Assessment: Summary for Policymakers. Figures, tables and captions in French

<p>French translations of the figures, tables and their captions from the Summary for Policymakers of the Thematic Assessment Report on Invasive Alien Species and their Control of the Intergovernmental Science-Policy Platform on Biodiversity and Ecosystem Services.</p>

opencc-by-4.0May 2024View details →
zenodo40/100

IPBES Invasive Alien Species Assessment: Summary for Policymakers. Figures, tables and captions in Russian

<p>Russian translations of the figures, tables and their captions from the Summary for Policymakers of the Thematic Assessment Report on Invasive Alien Species and their Control of the Intergovernmental Science-Policy Platform on Biodiversity and Ecosystem Services.</p>

opencc-by-4.0May 2024View details →
zenodo40/100

Keyframe-level Captions for V3C

<p>Image captions generated using <a href="https://github.com/j-min/CLIP-Caption-Reward">CLIP-Caption-Reward</a> and <a href="https://github.com/rmokady/CLIP_prefix_caption">CLIP_prefix_caption</a> for every Keyframe in the Vimeo Creative Commons Collection. Captions are grouped by captioning method and dataset shard and provided as CSV.</p>

opencc-by-4.0Nov 2022View details →
zenodo40/100

Charades-STA Speech Caption Dataset

<ul> <li><strong>Dataset introduction: </strong>This dataset is an extension of Charades-STA dataset, where audio is read from text using machine simulation method "microsoft/speecht5_tts"</li> <li><strong>Associated Code: <a href="https://github.com/xian-sh/UniSDNet">https://github.com/xian-sh/UniSDNet</a></strong></li> <li><strong>Associated Paper: <a href="https://arxiv.org/abs/2403.14174">https://arxiv.org/abs/2403.14174</a><br></strong></li> <li><strong>Disclaimer: </strong>This dataset is for academic research only, non-commercial use, if you use this dataset please cite the <a href="https://arxiv.org/abs/2403.14174">associated paper</a></li> <li><strong>Cite:</strong></li> </ul> <p>@article{hu2024unified,<br>&nbsp; title={Unified Static and Dynamic Network: Efficient Temporal Filtering for Video Grounding},<br>&nbsp; author={Jingjing Hu and Dan Guo and Kun Li and Zhan Si and Xun Yang and Xiaojun Chang and Meng Wang},<br>&nbsp; year={2024},<br>&nbsp; Journal={CoRR},<br>&nbsp; volume={abs/2403.14174},<br>}</p>

openJun 2023View details →
zenodo40/100

TACoS Speech Caption Dataset

<ul> <li><strong>Dataset Introduction: </strong>This dataset is an extension of Charades-STA dataset, where audio is read from text using machine simulation method "microsoft/speecht5_tts"</li> <li><strong>Associated Code: <a href="https://github.com/xian-sh/UniSDNet">https://github.com/xian-sh/UniSDNet</a></strong></li> <li><strong>Associated Paper: <a href="https://arxiv.org/abs/2403.14174">https://arxiv.org/abs/2403.14174</a><br></strong></li> <li><strong>Disclaimer: </strong>This dataset is for academic research only, non-commercial use, if you use this dataset please cite the <a href="https://arxiv.org/abs/2403.14174">associated paper</a></li> <li><strong>Cite:</strong></li> </ul> <p>@article{hu2024unified,<br>&nbsp; title={Unified Static and Dynamic Network: Efficient Temporal Filtering for Video Grounding},<br>&nbsp; author={Jingjing Hu and Dan Guo and Kun Li and Zhan Si and Xun Yang and Xiaojun Chang and Meng Wang},<br>&nbsp; year={2024},<br>&nbsp; Journal={CoRR},<br>&nbsp; volume={abs/2403.14174},<br>}</p>

openJun 2023View details →
zenodo40/100

CAPTDURE: Captioned Sound Dataset of Single Sources

<p><strong>Description</strong></p> <p>This&nbsp;is a dataset with captions for a single-source sound that can be used in various tasks that use environmental sounds. The dataset consists of 1,044 single-source sounds&nbsp;and 4,902 captions (3 or more captions per single-source sound). This dataset also consists of 1,044 multiple-source sounds and 3,132 captions (3 captions per multiple-source sound). The detail of the dataset is described in [1].</p> <p><strong>Conditions of use</strong></p> <p>This dataset was made by&nbsp;<strong>Hitachi, Ltd.</strong>&nbsp;and is available&nbsp;under a Creative Commons Attribution-NonCommercial-ShareAlike 4.0 International (CC BY-NC-SA 4.0) license.</p> <p><strong>Citation</strong></p> <p>If you use this dataset, please cite as follow:</p> <p>[1]&nbsp;Yuki Okamoto, Kanta Shimonishi, Keisuke Imoto, Kota Dohi, Shota Horiguchi, and Yohei Kawaguchi, &quot;CAPTDURE: Captioned sound Dataset of Single Sources,&quot; Proc. INTERSPEECH, pp. 1683-1687, 2023.</p> <p><strong>Feedback</strong></p> <p>If there is any problem, please contact us</p> <ul> <li>Yuki Okamoto, <a href="mailto:y-okamoto@ieee.org">y-okamoto@ieee.org</a></li> <li>Yohei Kawaguchi,&nbsp;<a href="mailto:yohei.kawaguchi.xk@hitachi.com">yohei.kawaguchi.xk@hitachi.com</a></li> </ul>

opencc-by-4.0Aug 2023View details →
zenodo36/100

IPBES Invasive Alien Species Assessment: Chapter 1. Figures, tables and captions

<p>Figures, tables and captions from Chapter 1: Introducing biological invasions and the IPBES thematic assessment of invasive alien species and their control. In: Thematic Assessment Report on Invasive Alien Species and their Control of the Intergovernmental Science-Policy Platform on Biodiversity and Ecosystem Services.</p>

openOct 2023View details →
zenodo36/100

IPBES Invasive Alien Species Assessment: Chapter 2. Figures, tables, captions and data management reports

<p>This folder contains the figures and tables included in Chapter 2 of the IPBES Invasive Alien Species and their Control Assessment Report. Each figure is provided in pdf and svg format. In addition, the R scripts and the data sets required to generate the figures are provided as well.</p><p>The figures and tables showing information about alien species numbers or distributions are all based on two data sets, which are stored on Zenodo folders. One data set contains the records of alien species per region worldwide (https://doi.org/10.5281/zenodo.7554428) and the workflow including R scripts have been published (https://doi.org/10.3897/neobiota.59.53578). This data set is called the 'chapter database'. Version 2.4.1 of this data set was used to extract the numbers shown in figures and tables of this chapter.&nbsp; The second data set (https://doi.org/10.5281/zenodo.6458083) contains coordinates of alien species occurrences worldwide and the workflow including R scripts have also been published elsewhere (https://doi.org/10.3897/neobiota.74.81082). Version 1.0.1 of this data set was used here.</p><p>The folder also contains the data management reports for the generation of the chapter database and figures.</p>

opencc-by-4.0Oct 2023View details →
zenodo36/100

Abstractive News Captions with High- level cOntext Representation (ANCHOR) dataset

<p><span>The</span> <span>Abstractive News Captions with High-</span><span>level cOntext Representation</span> <span>(</span><span>ANCHOR</span><span>) dataset contains 70K+ samples </span><span>sourced from 5 different news media organizations. This dataset can be utilized for Vision &amp; Language tasks such as Text-to-Image Generation, Image Caption Generation, etc.</span></p>

opencc-by-nc-4.0Apr 2024View details →
zenodo36/100

Places Audio Captions (Japanese) 100k

<p>The Places Audio Caption (Japanese) 100K Corpus contains approximately 100,000 Japanese&nbsp;spoken captions for natural images drawn from the Places 205 image dataset.</p> <p>This&nbsp;speech&nbsp;corpus&nbsp;was&nbsp;collected to investigate the learning of spoken language (words, sub-word units, higher-level semantics, etc.) from visually-grounded speech.&nbsp;For a description of the corpus, see:</p> <pre><code>@INPROCEEDINGS{Ohishi2020trilingual, author={Ohishi, Yasunori and Kimura, Akisato and Kawanishi, Takahito and Kashino, Kunio and Harwath, David and Glass, James}, booktitle={ICASSP 2020 - 2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)}, title={Trilingual Semantic Embeddings of Visually Grounded Speech with Self-Attention Mechanisms}, year={2020}, pages={4352-4356}, }</code></pre> <p>The&nbsp;corpus only includes audio recordings, and not the associated images. You will need to separately download the Places image dataset&nbsp;<a href="http://places.csail.mit.edu/">here</a>.</p> <p>The data&nbsp;is&nbsp;distributed under the Creative Commons Attribution-ShareAlike (CC BY-SA) license&nbsp;<a href="https://creativecommons.org/licenses/by-sa/4.0/legalcode">(link)</a>.</p> <p>If you use this data in your own publications, please cite the paper above.</p>

opencc-by-4.0Nov 2021View details →
zenodo36/100

Audio Caption Dataset (Hospital & Car)

<p>This dataset consists of the Hospital scene of our Audio Caption dataset. Details can be seen in our paper&nbsp;<a href="https://arxiv.org/abs/1902.09254">Audio Caption: Listen and Tell</a> published at ICASSP2019.&nbsp;</p> <p>Car scene, detailed in&nbsp;<a href="https://arxiv.org/abs/1905.13448">Audio Caption in a Car Setting with a Sentence-Level Loss</a>&nbsp;published at ISCSLP 2021.</p> <p>Original captions in Mandarin Chinese, with English translations provided.&nbsp;</p>

opencc-by-4.0May 2019View details →
zenodo36/100

Datasets related to the paper " Inception Models for Fashion Image Captioning: An Extensive Study on Multiple Datasets"

<p>This collection contains three datasets in HDF5 format: FashionCap, ReducedInFashAI, ReducedFACAD.</p>

opencc-by-4.0Oct 2022View details →
zenodo36/100

Captions of Vat. gr. 752 from the illustrations on Ps 1–151 – Dataset

<p>This is a beta-version.</p>

opencc-by-nc-4.0May 2024View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record