Skip to main content
zenodorestricted

Anonymized Dataset for "How Difficult Can It Be? Assessing Workload Components and Predictors in Introductory CS Projects"

<p>Paper title: How Difficult Can It Be? Assessing Workload Components and Predictors in Introductory CS Projects</p> <p>This is a replication package using the anonymized dataset that we have collected including 161 entries of NASA-TLX surveys and the associated student source code metrics. &nbsp;Please cite the paper properly if you plan on using any parts of the dataset or the replication package.</p> <p>&nbsp;&nbsp;<br> &nbsp;&nbsp;<br> The following preprocessing steps proceed the work in this file: &nbsp;<br> 1- Collecting survey data and filtering incomplete entries, anonymizing data by replacing names with numbers. &nbsp;<br> 2- Calculating TLX and adding it as a feature to the data. &nbsp;<br> 3- Processing the code associated with each entry to calculate: loc, functions/methods, classes, style errors. &nbsp;<br> 4- Processing submission system logs to calculate number of commits/submissions. &nbsp;<br> &nbsp;&nbsp;<br> &nbsp;&nbsp;</p> <p><br> Dataset: anonymous-complete-preprocessed-TLX.csv &nbsp;<br> key: anonymous-dataset-key.txt</p>

ShareScore

8/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
4
Harmonization
4
Access
0
Reuse readiness
0
Engagement
0