zenodoopen
Gaussian synthetic cluster datasets
<p>A collection of 20 structurally diverse synthetic datasets that consist of randomly generated gaussian distributions varying in number of objects (5000 or 10000), number of features (20,40,50,60), number of clusters (3,8,15,20), cluster sizes, cluster standard deviations, cluster overlap, and cluster anisotropy. Can be used to test clustering methods.</p>
ShareScore
32/100
Overall dataset sharing score
Score breakdown
These five areas show where the dataset supports — or may limit — practical reuse.
- Stewardship
- 12
- Harmonization
- 4
- Access
- 16
- Reuse readiness
- 0
- Engagement
- 0