Skip to main content
zenodoopen

Data of LLMs evaluation

<p>This data is the outcome of evaluating the performance of three LLMs, including gpt-4o, gemini-1.5-pro, and claude-sonnet, on five tasks related to our system's functionalities, including generating corresponding pronunciation, example, synonyms, antonyms and contextual explanation.</p>

ShareScore

32/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
4
Harmonization
4
Access
16
Reuse readiness
8
Engagement
0