Skip to main content
zenodoopen

DUAL-T User Testing Dataset

<p><em><strong>The file</strong></em></p> <p>This file contains keylogging data from 21 professional literary translators, who translated three short stories in three different workflows. Data was collected between October 2023 and January 2024.</p> <p>&nbsp;</p> <p><em><strong>Participants</strong></em></p> <p>Participants have been assigned codes from <strong>P01 to P23</strong>. Originally, there were 24 participants, however logs for P04a, P09, and P20 had to be excluded from analysis due to technical errors, so they are not present in the spreadsheet.</p> <p>&nbsp;</p> <p><em><strong>Workflows</strong></em></p> <p>The workflows were:</p> <p><strong>WF01</strong>: Microsoft Word</p> <p><strong>WF02</strong>: Trados Studio 2022</p> <p><strong>WF03</strong>: Machine-Translation Postediting Platform (proprietary tool, not available on the market)</p> <p>Participants had access to an internet browser (Chrome) for all three conditions.</p> <p>&nbsp;</p> <p><em><strong>Source texts</strong></em></p> <p>The texts used for the experiments were taken from the short story collection&nbsp;<em>One More Thing</em>, by BJ Novak (2014). The three short stories used are:<br><br><strong>T01</strong>: Rome</p> <p><strong>T02</strong>: The Beautiful Girl in the Bookstore</p> <p><strong>T03</strong>: They Kept Driving Faster and They Outrun the Rain</p> <p>&nbsp;</p> <p><strong><em>The keylogging data</em></strong></p> <p>Keylogging data was collected using Inputlog 8.0.0.17 (<a href="https://www.inputlog.net/">https://www.inputlog.net/</a>)</p> <p>The data available in the spreadsheet includes:</p> <ul> <li>Translation time (hh:mm:ss)</li> <li>Translation time (s)</li> <li>Translation time (m)</li> <li>Total&nbsp;<em>n</em> keystrokes</li> <li>Total&nbsp;<em>n</em> mouse actions</li> <li>Total&nbsp;<em>n</em> pauses</li> <li>Pause duration (s)</li> <li>Mean duration of pauses (s)</li> <li>Pause ratio</li> <li>Time spent inside tool (s)</li> <li>Time spent outside tool (s)</li> <li>Source text&nbsp;<em>n</em> words</li> <li>Source text&nbsp;<em>n</em> characters</li> <li>Target text&nbsp;<em>n</em> words</li> <li>Target text&nbsp;<em>n</em> characters</li> <li>Seconds per source text character</li> </ul>

ShareScore

44/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
8
Harmonization
4
Access
20
Reuse readiness
8
Engagement
4

Topics