zenodoopen
Characterizing Distributed Machine Learning Workloads on Apache Spark
<p>This dataset was used for our submission at Middleware'22 titled: "Characterizing Distributed ML Workloads"</p> <p>It will contains the description and the raw data, its format, as well as a detailed description of the cluster deployments used by these experiments.<br> </p> <p>The full paper is available here:</p> <p>https://dl.acm.org/doi/10.1145/3590140.3629112</p>
ShareScore
32/100
Overall dataset sharing score
Score breakdown
These five areas show where the dataset supports — or may limit — practical reuse.
- Stewardship
- 8
- Harmonization
- 4
- Access
- 8
- Reuse readiness
- 8
- Engagement
- 4