Skip to main content
zenodoopen

Synthetic gaussian datasets

<p>A collection of 16 structurally diverse synthetic datasets that consist of randomly generated gaussian distributions varying in number of objects (5000 or 10000), number of features (20,40,50,60), number of clusters (3,8,15,20), cluster sizes, cluster standard deviations, cluster overlap, and feature anisotropy. Can be used to test clustering methods.</p>

ShareScore

36/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
4
Harmonization
4
Access
20
Reuse readiness
8
Engagement
0