Skip to main content
zenodoopen

Learning Lenient Parsing & Typing via Indirect Supervision

<p><strong>Validation Set:</strong> We have shared 3 CSV files containing human-annotated validation sets of our paper (<em>Validation Data.zip</em>).</p> <p><strong>AST and Student Code Correction:</strong>&nbsp; For generating AST and code correction, please check the two files, AST.py and Top1.py ( in <em>AST &amp; Top-1.zip</em> ). In AST.py, we present the output of different parts of the program with an example. Please read that one before Top1.py. We follow the implementation of<a href="https://github.com/Lsdefine/attention-is-all-you-need-keras?fbclid=IwAR2UCy9NvD_PLHFVqM6B1VqvMVZzWRIS25BTG4nJudEhs3684RSOpGBhOyk"> https://github.com/Lsdefine/attention-is-all-you-need-keras</a>. Please check the remaining code in the above link. We made a minor correction in the dataloader.py to use two separate vocabulary cutoffs for input and output. dataloader1.py, transformer1.py, etc. are an exact replication of dataloader.py and transformer.py. Since we are using two models, we did it that way to avoid any conflict.</p> <p><strong>TypeFix:</strong> Check the code in <em>TypeFix.zip.</em></p> <p><em>We have also published data set for FragFix and BlockFix.</em></p>

ShareScore

32/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
4
Harmonization
4
Access
16
Reuse readiness
8
Engagement
0