Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
183
datasets available to search
ShareScore release 0.9.0
Dataset results
183 results for “Retraction”
Retractions in Humanities and Social Sciences: A Study of Retracted Papers from China
<p>This dataset presents a comprehensive study of retractions in the field of Humanities and Social Sciences (HSS), focusing specifically on retracted papers originating from China. </p>
retracted and commented papers related to non-coding RNA
<p>This is the dataset related to the research integrity analysis paper of non-coding RNA.</p> <p>The retracted and commented papers related to non-coding RNA and their publication years and citations are presented. It have the following sheets:</p> <ul> <li>Figure 1 <ul> <li>number of publications, retraction and comments of each year</li> </ul> </li> <li>Figure 2 <ul> <li>Percentage of retractions and Percentage of papers for major Prefixes/Keywords</li> </ul> </li> <li>retraction <ul> <li>retracted papers with their citation and related research integrity issue</li> </ul> </li> <li>comment <ul> <li>commented papers with thier citation and pubpeer url</li> </ul> </li> </ul>
A Dataset of Metadata of Articles Citing Retracted Articles
<p>This dataset comprises of metada of articles citing retracted publications. Originally, we obtained the DOIs from the<a href="https://www.irit.fr/~Guillaume.Cabanac/problematic-paper-screener"> Feet of Clay Detector of the Problematic Paper Screener (PPS - FoCD)</a>. Additional columns that were not provided in PPS were added using Crossref & Retraction Watch Database (CRxRW) and Dimensions API services. This detector flags publications that cite retracted articles with additional metadata. </p> <p>By querying the Dimensions API with the DOIs of the FoC articles, we acquired information such as more detailed document types (editorial, review article, research article), open access status (we only kept open access FoC articles in the dataset since we want to access the full-texts in the future), and research fields (classified according to the Australian and New Zealand Standard Research Classification (ANZSRC) Fields of Research (FoR), comprising of 23 main fields such as biological sciences, education.</p> <p>To get further information about the cited retracted articles in the dataset, we used the joint release of CRxRW. Using this dataset, we added the retraction reasons and retraction years. </p> <p>The original dataset was obtained from the PPS FoCD in December 2023. At this time there were 22558 total articles flagged in FoCD. Using the data filtering feature in PPS, we had a preliminary selection before downloading the first version of the dataset. We applied a filter to obtain:</p> <ul> <li>non-retracted citing articles at the time of data curation*</li> <li>open-access citing articles since we need the whole text to go forward with natural language processing tasks</li> <li>cited retracted articles with at least one scientific content related reason of retraction </li> <li>only articles (not monographs, chapters) to retain a unified text type</li> </ul> <p>More information about the usage of this dataset will be updated. </p> <p>*Current retraction status of the citing articles can be different since this is a static dataset and scientific literature is dynamic. </p>
Inputs and results of "A qualitative and quantitative analysis of open citations to retracted articles: the Wakefield 1998 et al.'s case"
<p>This repository contains the datasets and visualizations generated in our work: <strong>"A qualitative and quantitative analysis of open citations to retracted articles: the Wakefield 1998 et al.’s case"</strong>.</p> <p><strong>Note:</strong> the data are all contained inside the <strong><em>data.zip</em> </strong>file. You need to unzip the container to get access to all the files and directories listed below.</p> <p>The data (citations) gathered accompanied by their annotated characteristics are stored in <strong><em>data/</em>:</strong></p> <ul> <li><em><strong>"cits_features.csv": </strong></em>a dataset containing all the entities (rows in the CSV) which have cited the Wakefield et al. retracted article, and a set of features characterizing each citing entity (columns in the CSV). The features included are: DOI ("doi"), year of publication ("year"), the title ("title"), the venue identifier ("source_id"), the title of the venue ("source_title"), yes/no value in case the entity is retracted as well ("retracted"), the subject area ("area"), the subject category ("category"), the sections of the in-text citations ("intext_citation.section"), the value of the reference pointer ("intext_citation.pointer"), the in-text citation function ("intext_citation.intent"), the in-text citation perceived sentiment ("intext_citation.sentiment"), and a yes/no value to denote whether the in-text citation context mentions the retraction of the cited entity ("intext_citation.section.ret_mention").<br> <strong>Note: </strong>this dataset is licensed under a <a href="https://creativecommons.org/publicdomain/zero/1.0/legalcode">Creative Commons public domain dedication (CC0)</a>.</li> <li><em><strong>"cits_text.csv": </strong>this dataset stores the abstract ("abstract") and the in-text citations context ("intext_citation.context") </em>for each citing entity identified using the DOI value ("doi").<br> <strong>Note: </strong>the data keep their original license (the one provided by their publisher). This dataset is provided in order to favor the reproducibility of the results obtained in our work.</li> </ul> <p><strong>Topic modeling</strong></p> <p>We run a topic modeling analysis on the textual features gathered (i.e. abstracts and citation contexts). The results are stored inside the <em><strong>topic_modeling/</strong></em> directory. The topic modeling has been done using MITAO, a tool for mashing up automatic text analysis tools and creating a completely customizable visual workflow [1]. The topic modeling results for each textual feature are separated into two different folders, <em><strong>abstract/</strong></em> for the abstracts, and <em><strong>intext_cit/</strong></em> for the in-text citation contexts. Both the directories contain the datasets and visualizations generated using MITAO. </p> <p> </p> <p><strong>References</strong></p> <p>[1] Ferri, P., Heibi, I., Pareschi, L., & Peroni, S. (2020). MITAO: A User Friendly and Modular Software for Topic Modelling [JD]. PuntOorg International Journal, 5(2), 135–149. <a href="https://doi.org/10.19245/25.05.pij.5.2.3">https://doi.org/10.19245/25.05.pij.5.2.3</a></p>
Cranial kinesis facilitates quick retraction of stuck woodpecker beaks
<p class="MsoNormal"><span>Much like nails that are hammered into wood, the beaks of woodpeckers regularly get stuck upon impact. A kinematic video analysis of pecking by black woodpeckers shows how they manage to quickly withdraw their beaks, revealing a two-phase pattern: first a few degrees of nose-down rotation about the nasofrontal hinge causes the tip of the upper beak to be retruded while its proximal end is lifted. Next, the head is lifted, causing nose-up rotation about the nasofrontal hinge while the lower beak starts retruding and initiates the final freeing. We hypothesise that these consecutive actions, taking place in about 0.05 s, facilitate beak retraction by exploiting the presumably low frictional resistance between the upper and lower beak keratin surfaces, allowing them to slide past each other. It also demonstrates the counter-intuitive value of maintaining cranial kinesis in a species adapted to deliver forceful impacts.</span></p>
Inputs and results of "A quantitative and qualitative citation analysis to retracted articles in the humanities domain"
<p>This repository contains the datasets and visualizations generated in our work: <strong>"A quantitative and qualitative citation analysis to retracted articles in the humanities domain"</strong>.</p> <p><strong>Note:</strong> the data are all contained inside the <strong><em>data.zip</em> </strong>file. You need to unzip the container to get access to all the files and directories listed below.</p> <p>The data (citations) gathered accompanied by their annotated characteristics are stored in <strong><em>data/</em>:</strong></p> <ul> <li><em>cits.csv: </em>a dataset containing all the entities (rows in the CSV) which have cited a retracted article in the humanities domain. Each citing entity (row) is accompanied by a set of features (columns) that characterizes it.<br> <strong>Note: </strong>this dataset is licensed under a <a href="https://creativecommons.org/publicdomain/zero/1.0/legalcode">Creative Commons public domain dedication (CC0)</a>.</li> <li><em>content.csv: </em>a dataset containing the abstracts and the in-text citation contexts of all the citing entities gathered.<br> <strong>Note: </strong>the data keep their original license (the one provided by their publisher). This dataset is provided in order to favor the reproducibility of the results obtained in our work.</li> <li><em>excluded_hum_retractions.csv: </em>a list of the 12 humanities retracted articles with a humanities affinity score < 2, therefore excluded from the analysis. </li> </ul> <p> </p> <p><strong>Topic modeling</strong></p> <p>We run a topic modeling analysis on the textual features gathered (i.e. abstracts and citation contexts). The results are stored inside the <em><strong>topic_model/</strong></em> directory. The topic modeling has been done using MITAO, a tool for mashing up automatic text analysis tools and creating a completely customizable visual workflow [1]. The directory <em><strong>workflow/ </strong></em>contains the workflows used in MITAO. The topic modeling results for each textual feature are separated into two different folders, <em><strong>abstract/</strong></em> for the abstracts, and <em><strong>cits_context/</strong></em> for the in-text citation contexts. Both the directories contain the following directories/files: </p> <ul> <li> <p><em><strong>datasets_and_views/: </strong></em>the datasets and visualizations generated using MITAO. </p> </li> <li> <p><em><strong>ldamodel_corpus_dict/: </strong></em>it contains the dictionary, the LDA topic model, and the tokenized and vectorized corpus.</p> </li> <li><em><strong>rawdata/: </strong></em>the textual collection, metadata, and stopwords used as input in the workflow of MITAO</li> </ul> <p> </p> <p><strong>References</strong></p> <p>[1] Ferri, P., Heibi, I., Pareschi, L., & Peroni, S. (2020). MITAO: A User Friendly and Modular Software for Topic Modelling [JD]. PuntOorg International Journal, 5(2), 135–149. <a href="https://doi.org/10.19245/25.05.pij.5.2.3">https://doi.org/10.19245/25.05.pij.5.2.3</a></p> <p> </p> <ol> </ol>
THREE-DIMENSIONAL VOLUMETRIC EVALUATION OF ROOT RESORPTION IN MAXILLARY ANTERIORS FOLLOWING EN-MASSE RETRACTION WITH VARYING FORCE VECTORS - A RANDOMIZED CONTROL TRIAL
<p>To assess the severity of root resorption (RR) during retraction of maxillary anteriors with and without skeletal anchorage (three different force systems); and analyze it comprehensively using cone-beam computed tomography (CBCT) superimpositions.</p>
Methodology data of "A qualitative and quantitative citation analysis toward retracted articles: a case of study"
<p>This document contains the datasets and visualizations generated after the application of the methodology defined in our work: <em>"A qualitative and quantitative citation analysis toward retracted articles: a case of study"</em>. The methodology defines a citation analysis of the Wakefield et al. [1] retracted article from a quantitative and qualitative point of view. The data contained in this repository are based on the first two steps of the methodology. The first step of the methodology (i.e. “Data gathering”) builds an annotated dataset of the citing entities, this step is largely discussed also in [2]. The second step (i.e. "Topic Modelling") runs a topic modeling analysis on the textual features contained in the dataset generated by the first step. </p> <p><strong>Note:</strong> the data are all contained inside the "<strong><em>method_data.zip"</em> </strong>file. You need to unzip the file to get access to all the files and directories listed below.</p> <p> </p> <p><strong>Data gathering</strong></p> <p>The data generated by this step are stored in <strong>"<em>data/</em>"</strong>:</p> <ol> <li><em><strong>"cits_features.csv": </strong></em>a dataset containing all the entities (rows in the CSV) which have cited the Wakefield et al. retracted article, and a set of features characterizing each citing entity (columns in the CSV). The features included are: DOI ("doi"), year of publication ("year"), the title ("title"), the venue identifier ("source_id"), the title of the venue ("source_title"), yes/no value in case the entity is retracted as well ("retracted"), the subject area ("area"), the subject category ("category"), the sections of the in-text citations ("intext_citation.section"), the value of the reference pointer ("intext_citation.pointer"), the in-text citation function ("intext_citation.intent"), the in-text citation perceived sentiment ("intext_citation.sentiment"), and a yes/no value to denote whether the in-text citation context mentions the retraction of the cited entity ("intext_citation.section.ret_mention").<br> <strong>Note: </strong>this dataset is licensed under a <a href="https://creativecommons.org/publicdomain/zero/1.0/legalcode">Creative Commons public domain dedication (CC0)</a>.<br> </li> <li><em><strong>"cits_text.csv": </strong>this dataset stores the abstract ("abstract") and the in-text citations context ("intext_citation.context") </em>for each citing entity identified using the DOI value ("doi").<br> <strong>Note: </strong>the data keep their original license (the one provided by their publisher). This dataset is provided in order to favor the reproducibility of the results obtained in our work.</li> </ol> <p> </p> <p><strong>Topic modeling</strong><br> We run a topic modeling analysis on the textual features gathered (i.e. abstracts and citation contexts). The results are stored inside the <em><strong>"topic_modeling/"</strong></em> directory. The topic modeling has been done using MITAO, a tool for mashing up automatic text analysis tools, and creating a completely customizable visual workflow [3]. The topic modeling results for each textual feature are separated into two different folders, <em><strong>"abstracts/"</strong></em> for the abstracts, and <em><strong>"intext_cit/"</strong></em> for the in-text citation contexts. Both the directories contain the following directories/files: <strong> </strong></p> <ol> <li> <p><em><strong>"mitao_workflows/"</strong></em>: the workflows of MITAO. These are JSON files that could be reloaded in MITAO to reproduce the results following the same workflows.</p> </li> <li> <p><em><strong>"corpus_and_dictionary/": </strong></em>it contains the dictionary and the vectorized corpus given as inputs for the LDA topic modeling.</p> </li> <li> <p><em><strong>"coherence/coherence.csv":</strong></em> the coherence score of several topic models trained on a number of topics from 1 - 40.</p> </li> <li> <p><em><strong>"datasets_and_views/": </strong></em>the datasets and visualizations generated using MITAO. </p> </li> </ol> <p> </p> <p><strong>References</strong></p> <ol> <li>Wakefield, A., Murch, S., Anthony, A., Linnell, J., Casson, D., Malik, M., Berelowitz, M., Dhillon, A., Thomson, M., Harvey, P., Valentine, A., Davies, S., & Walker-Smith, J. (1998). RETRACTED: Ileal-lymphoid-nodular hyperplasia, non-specific colitis, and pervasive developmental disorder in children. <em>The Lancet</em>, <em>351</em>(9103), 637–641. <a href="https://doi.org/10.1016/S0140-6736(97)11096-0">https://doi.org/10.1016/S0140-6736(97)11096-0</a></li> <li> <p>Heibi, I., & Peroni, S. (2020). A methodology for gathering and annotating the raw-data/characteristics of the documents citing a retracted article v1 (protocols.io.bdc4i2yw) [Data set]. In protocols.io. ZappyLab, Inc. <a href="https://doi.org/10.17504/protocols.io.bdc4i2yw">https://doi.org/10.17504/protocols.io.bdc4i2yw</a></p> </li> <li> <p> </p> Ferri, P., Heibi, I., Pareschi, L., & Peroni, S. (2020). MITAO: A User Friendly and Modular Software for Topic Modelling [JD]. PuntOorg International Journal, 5(2), 135–149. <a href="https://doi.org/10.19245/25.05.pij.5.2.3">https://doi.org/10.19245/25.05.pij.5.2.3</a> <p> </p> </li> </ol>
VatsalSy/Taylor-Culick-retractions: Taylor-Culick retractions
This repository is under construction and will contain all the codes for reproducibility of the results in the paper: Taylor-Culick retraction at a liquid-air interface by Sanjay et al. (2021)
Data and results of "Retractions in Arts and Humanities: an analysis of the retraction notices"
<p>This repository contains the datasets and visualizations generated in our work: <strong>"Retractions in Arts and Humanities: an analysis of the retraction notices"</strong>.</p> <p><strong>Note:</strong> the data are all contained inside the <strong><em>data.zip</em> </strong>file. You need to unzip the container to get access to all the files and directories listed below.</p> <p><strong>Metadata</strong></p> <p>The directory <em><strong>metadata/ </strong></em>contains a CSV with the citation count of all the retracted papers we have considered. Metadata retrieved from Retraction Watch cannot be published in this repository due to copy rights issues. </p> <p><strong>Content analysis</strong></p> <p>We run a topic modeling analysis on the content of the retraction notices. The topic modeling analysis has been done using MITAO, a tool for mashing up automatic text analysis tools and creating a completely customizable visual workflow [1]. The topic modeling data and results are separated into the following directories/files: </p> <ul> <li> <p><em><strong>workflow/ </strong></em>contains the workflow used in MITAO.</p> </li> <li> <p><em><strong>datasets_and_views/: </strong></em>the datasets and visualizations generated using MITAO. </p> </li> <li> <p><em><strong>ldamodel_corpus_dict/: </strong></em>it contains the dictionary, the LDA topic model, and the tokenized and vectorized corpus.</p> </li> <li><em><strong>rawdata/: </strong></em>the textual collection, metadata, and stopwords used as input in the workflow of MITAO</li> </ul> <p><strong>References</strong></p> <p>[1] Ferri, P., Heibi, I., Pareschi, L., & Peroni, S. (2020). MITAO: A User Friendly and Modular Software for Topic Modelling [JD]. PuntOorg International Journal, 5(2), 135–149. <a href="https://doi.org/10.19245/25.05.pij.5.2.3">https://doi.org/10.19245/25.05.pij.5.2.3</a></p>
The Role of Pre-deployment Retraction in Decreasing Biopsy Clip Migration During Stereotactic Breast Biopsies
ClinicalTrials.gov study NCT04398537. IPD Sharing: NO. Countries: 1. Publications: 1.
Platelet-Rich Plasma (PRP) in Reconstructive Surgery on Children With Retractable Burn Sequelae on Extremities
ClinicalTrials.gov study NCT00858442. IPD Sharing: Not stated. Countries: 1. Publications: 6.
Cranial kinesis facilitates quick retraction of stuck woodpecker beaks
Open the record for dataset details and reuse information.
Figure 2 in Retraction to "Protocatechuic acid attenuates cerebral aneurysm formation and progression by inhibiting TNF-alpha/Nrf-2/NF-kB-mediated inflammatory mechanisms in experimental rats".
Figure 2: Lateral (above), dorsal (below, left) and ventral (below, right) views of the skull and labial view of the mandible (middle) of a male CtenomYS dorSaliS (FMNH 54391) from Colonia Fernheim, Boquerón, Paraguay. Scale = 5 mm.
Figure 3 in Retraction to "Protocatechuic acid attenuates cerebral aneurysm formation and progression by inhibiting TNF-alpha/Nrf-2/NF-kB-mediated inflammatory mechanisms in experimental rats".
Figure 3: Map of Paraguay and surroundings, showing the newly documented localities for CtenomYS dorSaliS (1 – Paraguay: Boqueron; Colonia Fernheim, 16 km W Filadelfia; and 2 – Paraguay: Boqueron; Colonia Mennonita, Orloff). Localities discussed by Contreras and Roig (1992) as possible type localities of this species include: 3 – Misión Inglesa, Waikhtlatingmalyalwa and 4 – a point between Las Juntas and Fortín Page. The dotted line corresponds to the limit between Bolivia and Paraguay according to the Treaty of Benítez-Ichazo (23th November 1894). The suggested distribution for this species proposed by Bidau (2015) is indicated in light blue (??). Landscape cover represents the ecoregions according to Olson et al. (2001).
Figure 1 in Retraction to "Protocatechuic acid attenuates cerebral aneurysm formation and progression by inhibiting TNF-alpha/Nrf-2/NF-kB-mediated inflammatory mechanisms in experimental rats".
Figure 1: Dorsal, ventral and lateral view of the skin of a female of CtenomYS dorSaliS (FMNH 54396) from Colonia Fernheim, Boquerón, Paraguay.
Figure 5 in Retraction to "Protocatechuic acid attenuates cerebral aneurysm formation and progression by inhibiting TNF-alpha/Nrf-2/NF-kB-mediated inflammatory mechanisms in experimental rats".
Figure 5: Pairwise distance values for divergence within the genus CtenomYS, including interspecific (range 0.029–0.077, with an average of 0.046) and intraspecific (range = 0–0.049, with an average of 0.008) ranges and values for divergence within C. dorSaliS and between C. dorSaliS and various other members of the genus.
Figure 4 in Retraction to "Protocatechuic acid attenuates cerebral aneurysm formation and progression by inhibiting TNF-alpha/Nrf-2/NF-kB-mediated inflammatory mechanisms in experimental rats".
Figure 4: Garli tree inferred from the maximum likelihood (ML) analysis of the 1140 bp Cytb dataset, partitioned by first, second and third codon position, under the GTR + G4 + I model for the three codon positions of sequence evolution. Node values represent Garli1000 bootstrap replicate percentages (BPs).
Data from: Cretaceous lophocoronids with short proboscis and retractable female genitalia provide the earliest evidence for their feeding and oviposition habits
<p>We describe two new species of Lophocoronidae: <em>Acanthocorona hedida</em> Zhang, Shih and Engel <strong>sp. n.</strong> and <em>Acanthocorona venulosa </em>Zhang, Shih and Engel <strong>sp. n.</strong>, and an undetermined specimen from mid-Cretaceous Kachin amber. Phylogenetic analysis of basal lepidopteran lineages, including three extinct families, was undertaken. The analysis supported monophyly of Glossata although internal relationships remain controversial. <em>Acanthocorona </em>and <em>Lophocorona </em>form a monophyletic group. It is likely that short and simply structured proboscides of <em>Acanthocorona </em>were used to sip water droplets, pollination drops from gymnosperms, nectar from early flowers, or sap from injured leaves. Both retracted and extended ovipositors are preserved in the material reported here, revealing their morphology and indicating that these Cretaceous lophocoronids inserted eggs into the tissues of their host plants.</p>
Extended dataset for retracted paper "Quantized Majorana Conductance"
<p>This repository contains an extended dataset of raw data underlying the retracted paper <a href="https://doi.org/10.1038/nature26142">H. Zhang et al, RETRACTED ARTICLE: Quantized Majorana conductance, Nature <strong>556</strong>, 74 (2018)</a></p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.