Automating Code Review Activities 2.0 (datasets and models)
<p>Resources related by the research work <em>"Automating Code Review Activities 2.0".</em></p> <ul> <li><strong>dataset.zip </strong>contains all the preprocessed datasets used in our work;</li> <li><strong>models.zip</strong> contains the (best) checkpoints of the fine-tuned T5 models;</li> <li><strong>tokenizer.zip</strong> contains the Sentencepiece model and vocabulary trained on our pre-training dataset;</li> <li><strong>automating_code_review.zip</strong> contains the material to successfully run our Colab notebooks.</li> </ul> <p>More information in the replication package of our work: <a href="https://github.com/CodeReviewAutomation/code_review_automation">code_review_autmoation</a></p>
ShareScore
8/100
Overall dataset sharing score
Score breakdown
These five areas show where the dataset supports — or may limit — practical reuse.
- Stewardship
- 4
- Harmonization
- 4
- Access
- 0
- Reuse readiness
- 0
- Engagement
- 0