zenodoopen
Data of LLMs evaluation
<p>This data is the outcome of evaluating the performance of three LLMs, including gpt-4o, gemini-1.5-pro, and claude-sonnet, on five tasks related to our system's functionalities, including generating corresponding pronunciation, example, synonyms, antonyms and contextual explanation.</p>
ShareScore
32/100
Overall dataset sharing score
Score breakdown
These five areas show where the dataset supports — or may limit — practical reuse.
- Stewardship
- 4
- Harmonization
- 4
- Access
- 16
- Reuse readiness
- 8
- Engagement
- 0