TACoS Speech Caption Dataset
<ul> <li><strong>Dataset Introduction: </strong>This dataset is an extension of Charades-STA dataset, where audio is read from text using machine simulation method "microsoft/speecht5_tts"</li> <li><strong>Associated Code: <a href="https://github.com/xian-sh/UniSDNet">https://github.com/xian-sh/UniSDNet</a></strong></li> <li><strong>Associated Paper: <a href="https://arxiv.org/abs/2403.14174">https://arxiv.org/abs/2403.14174</a><br></strong></li> <li><strong>Disclaimer: </strong>This dataset is for academic research only, non-commercial use, if you use this dataset please cite the <a href="https://arxiv.org/abs/2403.14174">associated paper</a></li> <li><strong>Cite:</strong></li> </ul> <p>@article{hu2024unified,<br> title={Unified Static and Dynamic Network: Efficient Temporal Filtering for Video Grounding},<br> author={Jingjing Hu and Dan Guo and Kun Li and Zhan Si and Xun Yang and Xiaojun Chang and Meng Wang},<br> year={2024},<br> Journal={CoRR},<br> volume={abs/2403.14174},<br>}</p>
ShareScore
40/100
Overall dataset sharing score
Score breakdown
These five areas show where the dataset supports — or may limit — practical reuse.
- Stewardship
- 8
- Harmonization
- 8
- Access
- 12
- Reuse readiness
- 8
- Engagement
- 4