Skip to main content
zenodoopen

CAPTDURE: Captioned Sound Dataset of Single Sources

<p><strong>Description</strong></p> <p>This&nbsp;is a dataset with captions for a single-source sound that can be used in various tasks that use environmental sounds. The dataset consists of 1,044 single-source sounds&nbsp;and 4,902 captions (3 or more captions per single-source sound). This dataset also consists of 1,044 multiple-source sounds and 3,132 captions (3 captions per multiple-source sound). The detail of the dataset is described in [1].</p> <p><strong>Conditions of use</strong></p> <p>This dataset was made by&nbsp;<strong>Hitachi, Ltd.</strong>&nbsp;and is available&nbsp;under a Creative Commons Attribution-NonCommercial-ShareAlike 4.0 International (CC BY-NC-SA 4.0) license.</p> <p><strong>Citation</strong></p> <p>If you use this dataset, please cite as follow:</p> <p>[1]&nbsp;Yuki Okamoto, Kanta Shimonishi, Keisuke Imoto, Kota Dohi, Shota Horiguchi, and Yohei Kawaguchi, &quot;CAPTDURE: Captioned sound Dataset of Single Sources,&quot; Proc. INTERSPEECH, pp. 1683-1687, 2023.</p> <p><strong>Feedback</strong></p> <p>If there is any problem, please contact us</p> <ul> <li>Yuki Okamoto, <a href="mailto:y-okamoto@ieee.org">y-okamoto@ieee.org</a></li> <li>Yohei Kawaguchi,&nbsp;<a href="mailto:yohei.kawaguchi.xk@hitachi.com">yohei.kawaguchi.xk@hitachi.com</a></li> </ul>

ShareScore

40/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
8
Harmonization
4
Access
20
Reuse readiness
8
Engagement
0

Topics