precomputed_data_HLA-EpiCheck
<p>This file contains precomputed data used in the HLA-Epicheck project (see https://gitlab.inria.fr/capsid.public_codes/hla-epicheck). This data can be re-used to run the notebook 'dataset_gen_al_radius_15.ipynb'.</p> <p>Data is organized by locus and antigen. In each antigen folder, four types of data can be found :</p> <ul> <li> <p>patchs : prepatches computed with the script compute_prepatches.tcl. Each file contains the prepatches for a given residue and patch radius (see file name). The format used in each file is as follows: each line corresponds to a prepatch and contains two colon-separated entries. The first one corresponds to a space-separated list of the residues that compose the prepatch and the second one corresponds to the PDB frame from which the prepatch was extracted.</p> </li> <li> <p>PDBs : PDB files used for computing the prepatches and SASA data.</p> </li> <li> <p>SASAs_out : SASA values computed for each PDB frame (i.e. one SASA file per frame). The format used in each file is as follows: each line corresponds to an AA and contains the AA number (numbering starts at 0) and the corresponding SASA value.</p> </li> <li> <p>RSASA_median.txt : median RSASA computed for each AA along the trajectory. The format used in each file is as follows: each line corresponds to an AA and contains the AA number (numbering starts at 0) and the corresponding RSASA value.</p> </li> </ul>
ShareScore
40/100
Overall dataset sharing score
Score breakdown
These five areas show where the dataset supports — or may limit — practical reuse.
- Stewardship
- 8
- Harmonization
- 4
- Access
- 16
- Reuse readiness
- 8
- Engagement
- 4