Skip to main content
zenodoopen

Simulated Cucurbit Tobamovirus Testing Dataset for Illustrating Possible Data Format Only

<p>This dataset contains wholly simulated and not particularly realistic data. It is intended to demonstrate the format of a possible dataset of tobamovirus testing results for cucurbit imports into Australia.</p> <p>If permission is obtained to publish real data, the real data will contain de-identified data on testing for virus presence in cucurbit seed lots where importation into Australia was attempted, over a period between 2019 and 2021. Each row represents a seed lot. For each lot, a number of groups of 400 seeds were tested for tobamovirus contamination, returning a presence or absence result for each group.There is no identifier or name representing the lot, consignment, importer or source country. The ID variable on the file was randomly assigned.</p> <p>Other authors will be added to the real data page subject to their permission and preference.</p> <p>For more information on the data, see&nbsp;</p> <p>Dall DJ, Lovelock DA, Penrose LDJ, Constable FE. Prevalences of Tobamovirus Contamination in Seed Lots of Tomato and Capsicum.&nbsp;<em>Viruses</em>. 2023; 15(4):883. https://doi.org/10.3390/v15040883</p> <p>&nbsp;</p> <p><strong>Variable Definitions</strong></p> <table> <tbody> <tr> <td> <p>Variable Name</p> </td> <td> <p>Values</p> </td> <td> <p>Description</p> </td> </tr> <tr> <td> <p>ID</p> </td> <td> <p>1, &hellip;, 800 (will be 793 in the real data)</p> </td> <td> <p>Randomised lot identifier number. The data set contains one row per lot.</p> </td> </tr> <tr> <td> <p>species</p> </td> <td> <p>tomato, capsicum</p> </td> <td> <p>species of cucurbit being imported</p> </td> </tr> <tr> <td> <p>year</p> </td> <td>2019, 2020, 2021</td> <td> <p>year in which importation was attempted</p> </td> </tr> <tr> <td> <p>month</p> </td> <td> <p>1, &hellip;, 12</p> </td> <td> <p>month in which importation was attempted</p> </td> </tr> <tr> <td> <p>NUM_TESTED</p> </td> <td> <p>1, ...., 50</p> </td> <td> <p>Number of groups tested. More groups were sampled from larger seed lots.</p> </td> </tr> <tr> <td>NUM_POS</td> <td>0, ..., 50</td> <td>Number of groups where virus contamination was detected.</td> </tr> <tr> <td> <p>vspecies</p> </td> <td> <p>PMMoV, PSTVd, PSTVd1, TASVd, TMV, ToBRFV, ToMMV, ToMV, blank</p> </td> <td> <p>Species of tobamovirus detected (equal to "blank", i.e., empty string, when no virus was detected.</p> </td> </tr> </tbody> </table> <p><strong>Author Field</strong></p> <p>If permission is obtained, this section will state that I am distributing this data with the permission of the data owner. (Other creators/contributers may be added to the "Creator" or "Contributer" section if they give permission for their names to appear.)</p>

ShareScore

36/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
4
Harmonization
4
Access
20
Reuse readiness
8
Engagement
0