Skip to main content
zenodorestricted

Raw data used to build models in SIMON Automated Machine Learning

<p>Here you can find all data and all information regarding each generated dataset.<br> For each dataset there are 4 files:</p> <p>json_info : This file contains, number of features with their names and number of subjects that are available for the same dataset<br> data_testing: data frame with data used to test trained model<br> data_training: data frame with data used to train models<br> results: direct unfiltered data from database</p> <p><br> Files are written in feather format.</p> <p><a href="https://gist.github.com/LogIN-/00d7628e0850f843ba84a678fac0a103">Here is an example</a> of data structure for each file in repository</p> <p>&nbsp;</p>

ShareScore

12/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
4
Harmonization
4
Access
0
Reuse readiness
0
Engagement
4