Skip to main content
zenodoopen

hadoop-14TB-part1

<p>Five data nodes worth of logs from a larger 14 TB dataset of Hadoop logs. The logs were&nbsp;generated from three Hadoop clusters, each containing 48 data nodes, running workloads from the HiBench Benchmark Suite for a month. This&nbsp;dataset was first used in the evaluation of &quot;CLP: Efficient and Scalable Search on Compressed Text Logs.&quot;</p>

ShareScore

36/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
4
Harmonization
8
Access
16
Reuse readiness
8
Engagement
0