Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

411

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

411 results for “arabic”

Learn how ShareScore rates datasets ↗
zenodo40/100

Arabic news credibility on Twitter using sentiment analysis and ensemble learning

<p>Arabic news credibility on Twitter using sentiment analysis and ensemble learning.</p> <p>&nbsp;</p> <p>WHAT IS IT?</p> <p>-----------</p> <p>an Arabic news credibility model on Twitter using sentiment analysis and ensemble learning.</p> <p>Here we include the Collected dataset and the source code of the proposed model written in Python language and using Keras library with Tensorflow backend.</p> <p>&nbsp;</p> <p>Required Packages</p> <p>------------------</p> <ol> <li>Keras (<a href="https://keras.io/">https://keras.io/</a>).</li> <li>Scikit-learn (<a href="http://scikit-learn.org/)">http://scikit-learn.org/)</a></li> <li>Imnlearn (<a href="https://imbalanced-learn.org/stable/">imbalanced-learn documentation &mdash; Version 0.10.1</a>)</li> </ol> <p>&nbsp;</p> <p>&nbsp;</p> <p>To Run the model</p> <p>---------------</p> <p>One data file is required to run the model which are:</p> <p>&nbsp;</p> <ol> <li>The data that were used are the collected dataset in the file, set the path of the required data file in the code.</li> </ol> <p>&nbsp;</p> <p>The dataset</p> <p>---------------</p> <ol> <li>There are the dataset file with all features, you can choose the features that you need and apply it on the model.</li> <li>There are a description file that describe each feature in the news credibility dataset</li> <li>The file Tweet_ID contains the list of tweets id in the dataset.</li> <li>The annotated replies based on credibility is provided.</li> </ol> <p>&nbsp;</p> <p>&nbsp;</p> <p>&nbsp;</p> <p>&nbsp;</p> <p>CONTACTS</p> <p>--------</p> <ul> <li>If you want to report bugs or have general queries email to &lt;duha_atif@yahoo.com&gt;</li> </ul> <p>&nbsp;</p> <p>&nbsp;</p> <p>&nbsp;</p>

opencc-by-4.0Jun 2023View details →
zenodo36/100

Figure 2 in Mapping the terrestrial reptile distributions in Oman and the United Arab Emirates

Figure 2. The Persian Wonder Gecko Teratoscincus keyserlingii photographed near Jebel Ali, Dubai.

opencc-by-4.0Dec 2009View details →
zenodo36/100

Fig. 3 in A new species of the genus Phytocoris (Heteroptera: Miridae) from the United Arab Emirates

Fig. 3. Phytocoris sweihanus sp. nov.: A – aedeagus; B-C – spiculum.

opencc-by-4.0Nov 2006View details →
zenodo36/100

Arab First Baptist Church

In what we refer to as The Great Commission, Jesus' command is clear. We are called to "Go, therefore, and make disciples of all nations, baptizing them in the name of the Father and of the Son and of the Holy Spirit, teaching them to observe everything I have commanded you." To be faithful to God's Word, we must take part in His mission to make disciples. Making disciples is the mission of Christ's church. -- www.arabfbc.org Source: Objaverse 1.0 / Sketchfab

opencc-byApr 2022View details →
zenodo36/100

Syrian Arab Republic

[Syria](https://en.wikipedia.org/wiki/Syria) is a country in Western Asia. It borders the Mediterranean Sea to the west, Turkey to the north, Iraq to the east and southeast, Jordan to the south, and Israel and Lebanon to the southwest. Cyprus lies to the west across the Mediterranean Sea. A country of fertile plains, high mountains, and deserts, Syria is home to diverse ethnic and religious groups, including the majority Syrian Arabs, Kurds, Turkmens, Assyrians, Armenians, Circassians, Mandaeans, and Greeks. Religious groups include Sunnis, Christians, Alawites, Druze, Isma'ilis, Mandaeans, Shiites, Salafis, and Yazidis. The capital and largest city of Syria is Damascus. Arabs are the largest ethnic group, and Sunnis are the largest religious group. Source: Objaverse 1.0 / Sketchfab

opencc-byFeb 2022View details →
zenodo36/100

DB_Expressive_Arabic

<p>This is a free database of expressive speech for the Arabic language. The database contains 13 speakers uttering a set of 10 sentences in 4 expressive styles (neutral, sad, joyful, and angry). The utterances of the first speaker were annotated using Praat. The files are in the mp3, ogg and wav formats&nbsp;</p>

opencc-zeroJul 2015View details →
zenodo36/100

United Arab Emirates

The [United Arab Emirates](https://en.wikipedia.org/wiki/United_Arab_Emirates) (UAE) is a country in Western Asia. It is located at the eastern end of the Arabian Peninsula, and shares borders with Oman and Saudi Arabia, while having maritime borders in the Persian Gulf with Qatar and Iran. Abu Dhabi is the nation's capital, while Dubai, the most populous city, is an international hub. Source: Objaverse 1.0 / Sketchfab

opencc-byFeb 2022View details →
zenodo36/100

Bangsa Arab Pra Islam

Mata Kuliah : Sejarah Peradaban Islam Materi : Bangsa Arab Pra Islam Objek : Ka'bah &amp; Berhala Menceritakan suasana bangsa arab dan mengenalkan bagian bagian Ka'bah. Source: Objaverse 1.0 / Sketchfab

opencc-byAug 2021View details →
zenodo36/100

Arabic Handwritten Legal Amount (AHLA) Dataset

<p>The AHLA dataset is collected by distributing an advanced designed report with Arabic native speakers. Our dataset contains two kinds of Arabic handwritten :&nbsp;<br>(1) &nbsp; &nbsp;Arabic word-level images that express legal amounts of bank cheques, including the colloquial words used in writing Arabic numbers.&nbsp;<br>(2) &nbsp; &nbsp;Arabic legal amount sentence images.</p> <p>The primary objective of compiling this comprehensive dataset is to furnish a diverse range of Arabic language samples. These samples are intended for training and testing systems capable of autonomously recognizing and comprehending handwritten legal amounts on financial documents. Subsequently, the aim is to convert these semantic expressions into their respective numeric currency totals, facilitating digital processing and banking operations.</p>

opencc-by-4.0Mar 2024View details →
zenodo36/100

Validation of the Arabic version of the mental toughness questionnaire

<p>The objective of this study was to assess the construct validity of the Arabic version of mental toughness MT based on the theoretical conception of the MTQ48 questionnaire <a href="#_ENREF_7">Clough et al. (2002)</a>. The 48 items consist of six components 6C :(1)Challenge CH, (2) Commitment CO, (3)Emotion Control EC, (4)Life Control LC, (5)Confidence in abilities CA,(6)Interpersonal confidence IC ,and the global score MT. The sample consisted of 853 Tunisian participants (444 males and 409 females; 409 athletes and 444 non-athletes), aged 14-27 years (<em>M=20.38 SD=4.12</em>). Cronbach&#39;s alpha suggests that over 48 items has adequate internal consistency (<em>&alpha;=.72)</em>. The EFA &nbsp;revealed a good sampling quality (df = 1128; p &lt; .001). The CFA approved a good model fit (<em>&chi;&sup2;=1146.33; df =1065; CFI=.93; SRMR=.063; RMSEA=.009</em>). In conclusion, our results allowed us to propose a valid Arabic measure of the 6C of MT. The results confirm the factorial validity of the MTQ48 and indicate that the Arabic version of the questionnaire has a robust measure of the psychometric properties of mental toughness. Finally, stakeholders in the Arab region should be able to benefit from the MTQ48 questionnaire, such as and reliable measurement and assessment instruments.</p>

opencc-by-4.0Aug 2021View details →
zenodo36/100

Training a Text-to-Speech System for Dialectal Arabic with a Focus on the Iraqi Dialect

<p>This research introduces a novel approach to Text-to-Speech (TTS) synthesis, focusing on the phonetic complexities of Arabic dialects, with particular emphasis on the Iraqi dialect. While existing Arabic speech corpora provide substantial coverage of Modern Standard Arabic (MSA), they fall short in capturing the phonetic richness of regional dialects. To address this gap, we utilized Nawar Halabi's Arabic Speech Corpus as a base dataset and enriched it with custom-recorded samples of the Iraqi dialect, incorporating distinctive phonemes such as گ ,ڤ ,پ ,چ ,ۆ ,ڵ ,ێ, and using the Tatweel character (ـ) as a vowel. Our approach, powered by the FastPitch model and a customized phonetiser, successfully synthesized the Iraqi dialect while also demonstrating adaptability to other Arabic dialects, including Egyptian, Khaliji, Syrian, and more. The results of this research signify a promising advancement in Arabic TTS technology, expanding its scope to authentically represent the diverse linguistic landscape of the Arabic-speaking world.</p>

opencc-by-4.0May 2024View details →
zenodo36/100

Dataset of Israeli-Jews, Palestinians, Israeli-Arabs, Americans and Cypriot students who played Fact Finders and PeaceMaker (examining game performance)

<p><strong>Serious Games, Players' Characteristics and Game Performance</strong></p> <p>The study examines the impact of players' gender, religiosity, political orientation, and nationality on game performance, focusing on two games for social impact revolving around intractable conflicts: PeaceMaker (Israeli-Palestinian conflict) and Fact Finders (Cyprus Conflict). Using questionnaires, we conducted a case study with 168 undergraduates voluntarily playing the two games and reporting their final scores. The participants included Israeli-Jew, Israeli-Arab, Palestinian, American, and Cypriot undergraduate students, differentiating between direct parties and third parties to the conflicts.</p>

opencc-by-4.0Jul 2024View details →
zenodo36/100

Arabic Poetry Analysis Datasets

<p>These datasets are used in the context of metered Arabic poetry analysis. They contain a poetry corpus, poems patterns data, poems visualizations as images and the data used for clustering poetry compositions with related results.</p>

opencc-by-4.0Dec 2023View details →
zenodo36/100

CLDF dataset derived from Ratcliffes's "Glottometrics of Arabic" from 2021

<p>Cite the source of the dataset as:</p> <blockquote> <p>Ratcliffe, Robert R. (2021): The glottometrics of Arabic. Language Dynamics and Change. 2021. DOI: 10.1163/22105832-01001100.</p> </blockquote>

opencc-by-4.0Jul 2021View details →
zenodo36/100

PAVONe: Platform of the Arabic Versions of the New Testament (pavone.uob-dh.org)

<p>PAVONe is a platform aiming at facilitating scholarly research on the early translations of the Gospels into Arabic and underlining the richness and diversity of these translations and their relations with the communities that have produced and used them. The platform is a database comprising a digital corpus of digitized and transcribed Arabic manuscripts of the four Gospels and lectionaries including both explicit and implicit verses of the Gospels with different layers of metadata (textual, paleographical, codicological, linguistic, etc.). In addition to this digital corpus, the platform provides a set of tools to enable scholars and researchers to manipulate those manuscripts and facilitate their study of the text.</p>

opencc-by-4.0May 2017View details →
zenodo36/100

Arab-Andalusian music corpus

<p>This repository contains Arab-Andalusian corpus collected in the CompMusic project.</p> <p>The following files are available for 164 concert recordings (overall playable time more than 125 hours):</p> <p>- Audio in mp3 format (44.1kHz or 48 kHz sampling, 128 Kbps and higher, mono or stereo)</p> <p>- Score in music xml format (manual transcriptions by the first author)</p> <p>- Automatically computed pitch (text format) and pitch distribution (json format) descriptors</p> <p>The meta data of the recordings (title, form, mizan, nawba and tab) are provided in separate json files. The corresponding MusicBrainz collection is available at <a href="https://musicbrainz.org/collection/142ea0d7-7fdf-4ea5-9b04-219f68023d01">this link</a>. Metadata is subject to improvements as it is collected via crowdsourcing on MusicBrainz. We gather and share&nbsp;new versions&nbsp;of the meta data (for the same audio content)&nbsp;<a href="https://zenodo.org/record/1299215#.WzOrb62B28o">at this link</a>.</p> <p>The lyrics for the recordings are available from <a href="https://zenodo.org/record/1291904#.Wyea5Bx9jCI">the Arab-Andalusian music lyrics dataset</a>.</p> <p>For more information, please refer to <a href="http://compmusic.upf.edu/corpora">http://compmusic.upf.edu/corpora</a></p> <p>A scientific publication making use of this database is available here:&nbsp;<a href="https://zenodo.org/record/1257388#.Wyeb9J9fjCI">Nawba Recognition for Arab-Andalusian Music Using Templates From Music Scores</a>.</p>

opencc-by-nc-4.0Jun 2018View details →
zenodo36/100

Figure 1 in Tabanidae Fauna (Order: Diptera) of the Arab Countries in the Middle East

Figure 1. Map showing the Arab countries in the Middle East.

opencc-by-4.0Jan 2024View details →
zenodo36/100

Table S1: Records of bramble shark sightings on 16 December, 2023, offshore of Fujairah, UAE, in the Gulf of Oman. Time is provided in Zulu time (GMT), not in local time. Vessels where sightings occurred are labeled as remotely operated vessels (ROV: Chimera) or human-driven submersibles (Sub: Neptune and Nadir).; Video S1: Video footage of a bramble shark (Echinorhinus brucus) encounter with a submersible at 780 m depth offshore from Fujairah, United Arab Emirates.

<p><span>Table S1: Records of bramble shark sightings on 16 December, 2023, offshore of Fujairah, UAE, in the Gulf of Oman. Time is provided in Zulu time (GMT), not in local time. Vessels where sightings occurred are labeled as remotely operated vessels (ROV: Chimera) or human-driven submersibles (Sub: Neptune and Nadir).</span></p> <p><span>Video S1: Video footage of a bramble shark (<em>Echinorhinus brucus</em>) encounter with a submersible at 780 m depth offshore from Fujairah, United Arab Emirates.</span></p>

opencc-by-4.0Sep 2024View details →
zenodo36/100

Arab Pangenome Reference

<p>Pangenomes represent a significant shift from relying on a single reference sequence to a robust set of assemblies, Arab populations remain significantly underrepresented; hence, we present the first Arab Pangenome Reference (APR) utilizing 53 individuals of diverse Arab ethnicities. We assembled nuclear and mitochondrial pangenomes using 35.27X high-fidelity long reads, 54.22X ultralong reads and 65.46X Hi-C reads yielded contiguous haplotype-phased&nbsp;<em>de novo</em> assemblies of exceptional quality, with an average N50 of 124.28 Mb. We discovered 111.96 million base pairs of novel euchromatic sequences absent from existing human pangenomes, the T2T-CHM13, GRCh38 reference human genomes, and other public datasets. We identified 8.94 million population-specific small variants and 235,195 structural variants within the Arab pangenome. We detected 883 gene duplications including 15.06% associated with recessive diseases and 1,436 bp of novel mitochondrial pangenome sequence. Our study provides a valuable resource for future genomic medicine initiatives in Arab population and other global populations.</p>

opencc-by-4.0Sep 2024View details →
zenodo36/100

Arabic Word Embedding Models

<p><strong>&nbsp;These&nbsp;are several Arabic Word Embedding Models for NLP tasks and it has been described in our paper titled &quot;</strong>Leveraging Arabic Sentiment Classification Using an Enhanced CNN-LSTM Approach and Effective Arabic Text Preparation<strong>&quot;&nbsp;&nbsp;</strong></p>

opencc-by-4.0Jun 2021View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record