BreXLiMe: A Semantically Enriched Dataset With News Articles, Micro-Posts, and TV Shows Related to the Brexit
<p>We provide a <strong>large data set of media content metadata</strong> from various media sources (including online news sites, social media, and live-TV) in three languages (<strong>English, German, and Spanish</strong>). Overall, the data set contains rich metadata for about <strong>240 thousand news articles, 12 million micro-posts, and 900 TV shows</strong>. All media content information has been semantically enriched with annotations of both entities and categories from DBpedia.</p> <p>The data can be used as a valuable data basis for applications and studies of various disciplines (e.g., social studies, political science, and humanities) on the case of Brexit, particularly on the <strong>media landscape before the Brexit referendum held on June 23, 2016</strong>.</p> <p>We provide the data set in the RDF serialization format Turtle (.ttl) as well as in XML.</p> <p>If you use our data set, please <strong>cite</strong> it as follows:</p> <pre><code>Lei Zhang, Maribel Acosta, Michael Färber, Steffen Thoma and Achim Rettinger. "BreXearch: Exploring Brexit Data Using Cross-Lingual and Cross-Media Semantic Search". In: Proceedings of the ISWC 2017 Posters & Demonstrations Track within the 16th International Semantic Web Conference (ISWC 2017). Vienna, Austria, 2017.</code></pre> <p> </p>
ShareScore
40/100
Overall dataset sharing score
Score breakdown
These five areas show where the dataset supports — or may limit — practical reuse.
- Stewardship
- 8
- Harmonization
- 8
- Access
- 16
- Reuse readiness
- 8
- Engagement
- 0