Skip to main content
zenodoopen

Gaussian synthetic cluster datasets

<p>A collection of 20 structurally diverse synthetic datasets that consist of randomly generated gaussian distributions varying in number of objects (5000 or 10000), number of features (20,40,50,60), number of clusters (3,8,15,20), cluster sizes, cluster standard deviations, cluster overlap, and cluster anisotropy. Can be used to test clustering methods.</p>

ShareScore

32/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
12
Harmonization
4
Access
16
Reuse readiness
0
Engagement
0

Topics