CAPTDURE: Captioned Sound Dataset of Single Sources
<p><strong>Description</strong></p> <p>This is a dataset with captions for a single-source sound that can be used in various tasks that use environmental sounds. The dataset consists of 1,044 single-source sounds and 4,902 captions (3 or more captions per single-source sound). This dataset also consists of 1,044 multiple-source sounds and 3,132 captions (3 captions per multiple-source sound). The detail of the dataset is described in [1].</p> <p><strong>Conditions of use</strong></p> <p>This dataset was made by <strong>Hitachi, Ltd.</strong> and is available under a Creative Commons Attribution-NonCommercial-ShareAlike 4.0 International (CC BY-NC-SA 4.0) license.</p> <p><strong>Citation</strong></p> <p>If you use this dataset, please cite as follow:</p> <p>[1] Yuki Okamoto, Kanta Shimonishi, Keisuke Imoto, Kota Dohi, Shota Horiguchi, and Yohei Kawaguchi, "CAPTDURE: Captioned sound Dataset of Single Sources," Proc. INTERSPEECH, pp. 1683-1687, 2023.</p> <p><strong>Feedback</strong></p> <p>If there is any problem, please contact us</p> <ul> <li>Yuki Okamoto, <a href="mailto:y-okamoto@ieee.org">y-okamoto@ieee.org</a></li> <li>Yohei Kawaguchi, <a href="mailto:yohei.kawaguchi.xk@hitachi.com">yohei.kawaguchi.xk@hitachi.com</a></li> </ul>
ShareScore
40/100
Overall dataset sharing score
Score breakdown
These five areas show where the dataset supports — or may limit — practical reuse.
- Stewardship
- 8
- Harmonization
- 4
- Access
- 20
- Reuse readiness
- 8
- Engagement
- 0