zenodoopen
hadoop-14TB-part1
<p>Five data nodes worth of logs from a larger 14 TB dataset of Hadoop logs. The logs were generated from three Hadoop clusters, each containing 48 data nodes, running workloads from the HiBench Benchmark Suite for a month. This dataset was first used in the evaluation of "CLP: Efficient and Scalable Search on Compressed Text Logs."</p>
ShareScore
36/100
Overall dataset sharing score
Score breakdown
These five areas show where the dataset supports — or may limit — practical reuse.
- Stewardship
- 4
- Harmonization
- 8
- Access
- 16
- Reuse readiness
- 8
- Engagement
- 0