Skip to main content
zenodoopen

Characterizing Distributed Machine Learning Workloads on Apache Spark

<p>This dataset was&nbsp;used for our submission at Middleware'22&nbsp;titled: "Characterizing Distributed ML Workloads"</p> <p>It will contains the description and the raw data, its format, as well as a detailed description of the cluster deployments used by these experiments.<br>&nbsp;</p> <p>The full paper is available here:</p> <p>https://dl.acm.org/doi/10.1145/3590140.3629112</p>

ShareScore

32/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
8
Harmonization
4
Access
8
Reuse readiness
8
Engagement
4

Topics