Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
22
datasets available to search
ShareScore release 0.9.0
Dataset results
22 results for “Serbian”
Wikipedia: wikipedia-sr (Serbian)
Wikipedia is a multilingual, web-based, free-content encyclopedia project supported by the Wikimedia Foundation and based on a model of openly editable content. EOL harvests articles from wikipedia that are indexed as species or higher taxa.<p></p>До сада је на Википедији на српском језику 252.241 корисник отворио налог, а од тога су 792 активна. Сви уредници су волонтери, који удружују напоре у оквиру различитих тематских целина. Дискусије и коментари садржаја чланака су добродошли. Странице за разговор се користе за размену мишљења и указивање на грешке, како би се постојећи чланци учинили што бољим и свеобухватнијим. <p></p>https://sr.wikipedia.org/wiki/
Climate Change Risk Assessment Dataset for Serbian Agriculture
<p>This dataset contains the assessment of observed climate change to the agriculture in Serbia. Data are given on a lat/lon gird with 0.01° horizontal resolution in geotif format. Climatic and bioclimatic indices, as well as risk occurrence frequencies are calculated using interpolated daily observations of temperature and precipitation over Serbia in the period 1998-2017. Statistical significance of the climate indices change, change in category of viticultural indices and increase in risk occurrence is given as well. Analyzed species are: vine grape, fruits (peaches, apricots, cherries, plums, apples, pears and quinces) and crops (corn, sunflower, sugar beet and winter wheat). Calculated climatic indices are: mean annual, vegetational and summer temperature and precipitation; bioclimatic indices: Winkler index, Huglin index, Dryness index, and Cool nights index; risks: drought, frost in the beginning of vegetation, low temeperature during winter, high temperature during summer and intensive rainfall.</p>
A multilabel dataset for distinguishing Bosnian, Croatian, Montenegrin, and Serbian
<p>This dataset contains files used in the VarDial 2024 Shared Task on Distinguishing Between Similar Languages - Multiple Labels for the Bosnian - Croatian - Montenegrin - Serbian (BCMS) subtask.</p> <p>The starting point for this dataset is the one published by Rupnik et al. (2023). It contains geolocated data from the BCMS linguistic area collected from Twitter (rebranded as X).<br>Each instance contains the full tweet production of a single user, which was manually annotated for the user's country.<br>The original annotation was single-label, and it was produced by a single annotator. In the version of the data produced here, the test and dev sets were reannotated by multiple annotators, in a multi-label setting. For the details on the reannotation process, please see Miletić and Miletić (2024). We have also excluded retweets from the original data, as these represent reproduced content from a different user account and may not be representative of the language use of the user themselves.</p> <p>For details on the shared task, we refer you to Chifu et al. (2024).</p>
Datasets for Crowd-based Requirements Engineering and aspect-based detection of learning-centered emotion from the text in Serbian language
<h1>Datasets for the paper "Enhancing Software and Learning with Serbian Student Feedback Corpora"</h1> <p>These datasets include student feedback on an Intelligent Tutoring System written in Serbian, annotated with categories for Crowd-based Requirements Engineering (CrowdRE) and aspect-based detection of learning-centered emotions. Four annotators manually annotated each sentence. </p> <p>The CrowdRE dataset includes two JSON files:</p> <ul> <li><strong>crowdre_english.json</strong> - annotated text with columns and classes written in English. Columns are: <ul> <li><em>Comment</em> - the entire student feedback</li> <li><em>Sentence</em> - sentence extracted from the feedback that was annotated</li> <li><em>Intention </em>- class representing the intention of the sentence</li> <li><em>Topic </em>- class representing the topic of the sentence</li> </ul> </li> <li><strong>crowdre_srpski.json</strong> - annotated text with columns and classes written in Serbian. Columns are:<br> <ul> <li><em>Komentar</em> - the entire student feedback</li> <li><em>Recenica </em>- sentence extracted from the entire feedback that was annotated</li> <li><em>Namera </em>- class representing the intention of the sentence</li> <li><em>Tema </em>- class representing the topic of the sentence.</li> </ul> </li> </ul> <p>The dataset for the aspect-based detection of learning-centered emotions includes two JSON files:</p> <ul> <li><strong>emotions_english.json </strong>- annotated text with columns and classes written in English. Columns are: <ul> <li><em>Comment </em>- the entire student feedback</li> <li><em>Sentence </em>- sentence extracted from the feedback that was annotated</li> <li><em>Aspect </em>- class representing the aspect of the sentence</li> <li><em>Emotion </em>- class representing the learning-centered emotion of the sentence</li> </ul> </li> <li><strong>emocije_srpski.json </strong>- annotated text with columns and classes written in Serbian. Columns are: <ul> <li><em>Komentar </em>- the entire student feedback</li> <li><em>Recenica</em> - sentence extracted from the entire feedback that was annotated</li> <li><em>Aspekt</em> - class representing the aspect of the sentence</li> <li><em>Emotion </em>- class representing the learning-centered emotion of the sentence.</li> </ul> </li> </ul> <p>Annotators annotated the dataset based on the annotation procedure and guidelines available <a href="https://github.com/Clean-CaDET/student-feedback-mining">here</a>. </p> <h2>Citation</h2> <p>If you use this in your research, please cite:</p> <blockquote> <p>Vidaković, D., Luburić, N., Kovačević, A., & Slivka, J. Enhancing software and learning with Serbian student feedback corpora. Language Resources & Evaluation (2025). https://doi.org/10.1007/s10579-025-09855-y</p> </blockquote> <p> </p>
Serbian Coronavirus News Comments Corpus News-CommSR
<p>A corpus of readers' news comments posted below news articles on the topic of the covid-19 pandemic, published in major Serbian daily newspapers and news portals in the six-month early pandemic period (March 2020 to September 2020).</p> <div>The corpus is designed to facilitate research on crisis discourses, crisis communication, as well as pandemic-time linguistic innovation. It is available in plain text version and XML with full metadata. The corpus complements a separate corpus of news articles Serbian Coronavirus Corpus NewsSR. Parallel versions from Croatia and Slovenia are also available.</div> <div> </div> <div>The project leading to this publication has received funding from the European Union’s Horizon 2020 research and innovation programme under the <a href="https://cordis.europa.eu/programme/id/H2020-EU.4./en">H2020-EU.4. - SPREADING EXCELLENCE AND WIDENING PARTICIPATION </a>programme Widening fellowships grant agreement No 101038047.</div>
Serbian Coronavirus Corpus NewsSR
<p>A corpus of news articles on the topic of the covid-19 pandemic, published in major Serbian daily newspapers and news portals in the six-month early pandemic period (March 2020 to September 2020).<br>The corpus is designed to facilitate research on crisis discourses, crisis communication, as well as pandemic-time linguistic innovation. It is available in plain text version and XML with full metadata. <br>Covid-NEWS-SR is complemented with a separate corpus of citizen metalanguage comments, i.e. online comments to the news articles, available as Covid-NEWS-Comm-SR. Parallel versions from Slovenia and Croatia are also available.</p> <p>The project leading to this publication has received funding from the European Union’s Horizon 2020 research and innovation programme under the <a href="https://cordis.europa.eu/programme/id/H2020-EU.4./en">H2020-EU.4. - SPREADING EXCELLENCE AND WIDENING PARTICIPATION </a>programme Widening fellowships grant agreement No 101038047.</p>
THE STATUS OF THE SERBIAN TERMINOLOGY DEFINED BY THE SERBIAN LANGUAGE POLICY THROUGHOUT ITS CONTEMPORARY AND FUTURE PLANS. AN OUTLINE OF ONE TERMINOLOGICAL ALGORITHM
<p>This lecture aims to point out the necessity of creating a digital terminology database of the Serbian professional terminology which would provide systematically collecting, inventory making, precise identifying, defining, linguistic analyzing and interpreting terminologies of different professions for making a stable ground for the codification and standardization of professional terminology of the Serbian language. It also implicitly points to a need to organize terminology workshops through which the experts that “build” vocabularies of their profession would become familiar with the basic principles of creating or designing terminology expressed in the work with the database via the network interface and thus become able to work on different professional terminologies. In addition, the paper suggests the most common mistakes in contemporary terminographic work as a result of lack of understanding of the meaning of inter- and multi-disciplinary approach to the study of terminology as a special branch of linguistic research.</p>
Serbian researchers papers in Web of Science database in the period 2003-2021
<p>This is dataset about Serbian papers published in journals indexed in Web of Science collections in the period 2003-2021<br> We used as "serbian" countries: Serbia, Serbia and Montenegro, Yugoslavia</p> <p>The terms Serbian researcher and Serbian paper used in this dataset are defined as follows:<br> Serbian researcher is a researcher affiliated with a Serbian institution,<br> Serbian paper is a paper with at least one Serbian researcher in the list of authors.<br> This means that a paper published by a researcher with non-Serbian nationality (i.e. German) working at a Serbian institution is taken into account in the analysis of Serbian papers described in this paper.</p> <p>Thomson Reuters’ Web of Science (WoS) database was used for data acquisition. We searched Serbian papers in two WoS collections - Science Citation Index Expanded (SCIE) and the Social Science Citation Index (SSCI). The search query executed over those collections on 18th of January, 2022.<br> CU=(%Serbia% OR %Serbia and Montenegro% OR %Yugoslavia%) AND PY=[2003-2021]<br> </p> <p>We also analyzed the number of articles of Serbian researchers published in journals with an unstable impact factor (IF), i.e. journals which didn’t have an IF before 2008, and had one in some subperiod of 2008-2015, i.e. lost their IF until 2015. The majority of those journals stopped being indexed in Web of Science as a ban for losing the quality or having predatory journal behavior. We found 143 such journals and their ISSNs by using the JCR (Journal citation reports in the period 2007 - 2015).<br> <br> </p>
PrisonLIFE project - Adaptation, translation equivalence and content validity of the MQPL survey in Serbian
<p>This dataset was created as part of the PrisonLIFE project, funded by the Science Fund of the Republic of Serbia under Grant No. 7750249.</p> <p>This dataset includes information about the adaptation process, translations, and content validity assessment of the Serbian version of the Measuring the Quality of Prison Life (MQPL) survey, along with details about participants in focus groups and their feedback. The study references the works by Liebling et al. (2012) and Milićević et al. (2023).</p>
Phase 3 Trial of Serbian Seasonal Influenza Vaccine
ClinicalTrials.gov study NCT02935192. IPD Sharing: NO. Countries: 1. Publications: 1.
FIGURES 8–11 in First records of Croatian and Serbian Tetrigidae (Orthoptera: Caelifera) with description of a new subspecies of Tetrix transsylvanica (Bazyluk & Kis, 1960)
FIGURES 8–11 (females). 8 and 9: head in front view: 8) T. t. hypsocorypha ssp. nov., 9) T. t. transsylvanica—9a) Cozia Mt., 9b) Timişu de Sus; 10 and 11) head and pronotum from above: 10) T. t. hypsocorypha ssp. nov., 11) T. t. transsylvanica—11a) Cozia Mt., 11b) Măgurii Cisnădiei, Sibiu. (8 and 10 photo Josip Skejo, 9a and 11a photo Sigfrid Ingrisch, 9b and 11b photo Ionuţ Ştefan Iorgu; reproduced with permission)
FIGURES 5–7. 5 in First records of Croatian and Serbian Tetrigidae (Orthoptera: Caelifera) with description of a new subspecies of Tetrix transsylvanica (Bazyluk & Kis, 1960)
FIGURES 5–7. 5) Holotype of Tetrix transsylvanica hypsocorypha Skejo, 2014 subspecies nova, general habitus. 6) Holotype labels. 7a.) Female of Tetrix transsylvanica transsylvanica from Cozia Mt.. 7b) Female of Tetrix transsylvanica transsylvanica from Timişu de Sus. (5, 6 photo Josip Skejo, 7a photo Sigfrid Ingrisch, 7b photo Ionuţ Ştefan Iorgu; reproduced with permission)
FIGURE 2 in First records of Croatian and Serbian Tetrigidae (Orthoptera: Caelifera) with description of a new subspecies of Tetrix transsylvanica (Bazyluk & Kis, 1960)
FIGURE 2. Male Tetrix tuerki from Franjo Košćec's collection. Photographed with Varaždin City Museum's (GMV) permission (photo: Josip Skejo).
FIGURE 1 in First records of Croatian and Serbian Tetrigidae (Orthoptera: Caelifera) with description of a new subspecies of Tetrix transsylvanica (Bazyluk & Kis, 1960)
FIGURE 1. Records of Tetrix undulata (triangles), and Tetrix tuerki (circles) in Croatia and Serbia.
FIGURES 12–17 in First records of Croatian and Serbian Tetrigidae (Orthoptera: Caelifera) with description of a new subspecies of Tetrix transsylvanica (Bazyluk & Kis, 1960)
FIGURES 12–17 (females). (Tetrix transsylvanica hypsocorypha Skejo, 2014 subspecies nova: 12, 13, 14, 16; Tetrix t. transsylvanica: 15, 17). 12) mesothorax, metathorax and abdomen (paratype). 13) tegmenulum and ala (paratype). 14) tegmenulum and ala from Nadig (1991). 15) tegmenulum and ala from Bazyluk & Kis (1960). 16) ovipositor (paratype). 17a) ovipositor (Timişu de Sus). 17b) ovipositor (Măgurii Cisnădiei) (12,13 and 16 photo Josip Skejo, 17a and 17b photo Ionuţ Ştefan Iorgu; reproduced with permission).
LittlEARS questionnaire on early speech production for the Serbian language in children with normal hearing
<p>This dataset includes data for validation of LittlEARS questionnaire on early speech production (LEESPQ) for the Serbian language in children with normal hearing.</p> <p>Dataset includes data about 206 children, age 0 to 18 months. Information was provided by their parents. </p>
Serbian Intangible Cultural Heritage in the Western Balkans: Perils and Prospects of Inclusive Research and Safeguarding (SICHWEB) project 1st year datasets
Open the record for dataset details and reuse information.
Psychometric properties of the World Health Organization's Quality of Life (WHOQOL-BREF) questionnaire in Serbian medical students
<p><strong>Psychometric properties of the World Health Organization's Quality of Life (WHOQOL-BREF) questionnaire in Serbian medical students </strong></p>
Effectiveness of Targeted Educational Intervention on Healthy Lifestyle Behaviors in Serbian Population
ClinicalTrials.gov study NCT02999425. IPD Sharing: YES. Countries: 1. Publications: 3.
Serbian Smoking Reduction/Cessation Trial (2SRT)
ClinicalTrials.gov study NCT00601042. IPD Sharing: Not stated. Countries: 1. Publications: 1.
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.