Skip to main content
zenodoopen

TACoS Speech Caption Dataset

<ul> <li><strong>Dataset Introduction: </strong>This dataset is an extension of Charades-STA dataset, where audio is read from text using machine simulation method "microsoft/speecht5_tts"</li> <li><strong>Associated Code: <a href="https://github.com/xian-sh/UniSDNet">https://github.com/xian-sh/UniSDNet</a></strong></li> <li><strong>Associated Paper: <a href="https://arxiv.org/abs/2403.14174">https://arxiv.org/abs/2403.14174</a><br></strong></li> <li><strong>Disclaimer: </strong>This dataset is for academic research only, non-commercial use, if you use this dataset please cite the <a href="https://arxiv.org/abs/2403.14174">associated paper</a></li> <li><strong>Cite:</strong></li> </ul> <p>@article{hu2024unified,<br>&nbsp; title={Unified Static and Dynamic Network: Efficient Temporal Filtering for Video Grounding},<br>&nbsp; author={Jingjing Hu and Dan Guo and Kun Li and Zhan Si and Xun Yang and Xiaojun Chang and Meng Wang},<br>&nbsp; year={2024},<br>&nbsp; Journal={CoRR},<br>&nbsp; volume={abs/2403.14174},<br>}</p>

ShareScore

40/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
8
Harmonization
8
Access
12
Reuse readiness
8
Engagement
4

Topics