zenodoopen
Sound-VECaps
<p>This is the dataset for Sound-VECaps, a large-scale audio dataset with visual-enhanced captions. </p> <p>We also release the dataset for AudioCaps-Enhanced, the visual-enhanced AudioCaps testing dataset as the new benchmark. </p>
ShareScore
40/100
Overall dataset sharing score
Score breakdown
These five areas show where the dataset supports — or may limit — practical reuse.
- Stewardship
- 4
- Harmonization
- 4
- Access
- 20
- Reuse readiness
- 8
- Engagement
- 4