Skip to main content
zenodoopen

Sound-VECaps

<p>This is the dataset for Sound-VECaps, a large-scale audio dataset with visual-enhanced captions.&nbsp;</p> <p>We also release the dataset for AudioCaps-Enhanced, the visual-enhanced AudioCaps testing dataset as the new benchmark.&nbsp;</p>

ShareScore

40/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
4
Harmonization
4
Access
20
Reuse readiness
8
Engagement
4