Skip to main content
zenodoopen

precomputed_data_HLA-EpiCheck

<p>This file contains precomputed data used in the HLA-Epicheck project (see https://gitlab.inria.fr/capsid.public_codes/hla-epicheck). This data can be re-used to run the notebook 'dataset_gen_al_radius_15.ipynb'.</p> <p>Data is organized by locus and antigen. In each antigen folder, four types of data can be found :</p> <ul> <li> <p>patchs : prepatches computed with the script compute_prepatches.tcl. Each file contains the prepatches for a given residue and patch radius (see file name). The format used in each file is as follows: each line corresponds to a prepatch and contains two colon-separated entries. The first one corresponds to a space-separated list of the residues that compose the prepatch and the second one corresponds to the PDB frame from which the prepatch was extracted.</p> </li> <li> <p>PDBs : PDB files used for computing the prepatches and SASA data.</p> </li> <li> <p>SASAs_out : SASA values computed for each PDB frame (i.e. one SASA file per frame). The format used in each file is as follows: each line corresponds to an AA and contains the AA number (numbering starts at 0) and the corresponding SASA value.</p> </li> <li> <p>RSASA_median.txt : median RSASA computed for each AA along the trajectory. The format used in each file is as follows: each line corresponds to an AA and contains the AA number (numbering starts at 0) and the corresponding RSASA value.</p> </li> </ul>

ShareScore

40/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
8
Harmonization
4
Access
16
Reuse readiness
8
Engagement
4

Topics