Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
2,609
datasets available to search
ShareScore release 0.9.0
Dataset results
2,609 results for “web”
Cothran, R. D., F. Radarian, and R. A. Relyea. 2011. Altering aquatic food webs with a global insecticide: Arthropod-amphibian links in mesocosms that simulate wetland communities. Journal of the North American Benthological Society 30:893-912.
Pesticides play a critical role in maximizing yields of economically important crops and minimizing the human health threats of disease-carrying pests, but they often have collateral effects on nontarget species. We used a mesocosm study to address how the most commonly used insecticide in the USA, malathion, applied at low, ecologically relevant concentrations (20 and 110 mg/L) affects species interactions in aquatic communities. Unlike many community ecotoxicology studies, our study assessed how malathion affects both consumptive and nonconsumptive effects of predators. We also considered how the vertical distribution of predator cues and malathion (caused by potential stratification) affects species interactions. We found no evidence for vertical stratification of malathion, a result suggesting that exposure to the pesticide was uniform throughout the water column. Malathion was lethal to some primary consumers (cladocerans) at both concentrations and to top predators (dragonflies) at the highest concentration (110 mg/L). These lethal effects initiated density-mediated indirect effects in both cases. Malathion also may have decreased dragonfly foraging efficiency, resulting in increased tadpole survival (trait-mediated indirect effect), which decreased the resources used by tadpoles (periphyton). Collectively, our results show that malathion alters species interactions. However, we suggest that the degree to which pesticides affect aquatic communities will depend strongly on the species composition of communities. Therefore, the community-level consequences of pesticide exposure are likely to vary across the ecological landscape.
Spider web distribution and characteristics in Dean Creek Marsh, Sapelo Island, Georgia, USA, October 2024
Salt marshes are a rare environment but are nonetheless home to many web-building spiders. This dataset describes a small-scale study of the distribution and sizes of webs in Dean Creek Marsh, located on the southern end of Sapelo Island, Georgia, USA. The study contained three components: A transect survey to understand the spatial distribution and density of spider webs among different vegetation types, a targeted search for webs to understand the population of webs in the region, and a sticky-trap study to investigate the prey abundance among vegetation types in the marsh.
SGS-LTER Long-Term Monitoring Project: Vegetation Cover on Small Mammal Trapping Webs on the Central Plains Experimental Range, Nunn, Colorado, USA 1999 -2006, ARS Study Number 118 (Reformatted to a Darwin Core Archive)
This data package is formatted as a Darwin Core Archive (DwC-A, event core). For more information on Darwin Core see https://www.tdwg.org/standards/dwc/. This Level 2 data package was derived from the Level 1 data package found here: https://pasta.lternet.edu/package/metadata/eml/edi/326/2, which was derived from the Level 0 data package found here: https://pasta.lternet.edu/package/metadata/eml/knb-lter-sgs/140/17. The abstract below was extracted from the Level 0 data package and is included for context: This data package was produced by researchers working on the Shortgrass Steppe Long Term Ecological Research (SGS-LTER) Project, administered at Colorado State University. Long-term datasets and background information (proposals, reports, photographs, etc.) on the SGS-LTER project are contained in a comprehensive project collection within the Digital Collections of Colorado (http://digitool.library.colostate.edu/R/?func=collections&collection_id=3429). The data table and associated metadata document, which is generated in Ecological Metadata Language, may be available through other repositories serving the ecological research community and represent components of the larger SGS-LTER project collection. Additional information and referenced materials can be found: http://hdl.handle.net/10217/83458. The abundance and diversity of small mammals in shortgrass steppe is strongly influenced by the structure and composition of vegetation. Vegetation structure provides cover from predators and harsh abiotic conditions. Plant species composition affects the types of seeds and herbaceous material available to granivores and herbivores, and influences arthropod populations, which are important prey for the omnivorous species that dominate in shortgrass steppe. Both vegetation structure and plant community composition are sensitive to the availability of precipitation as well as the activity of large mammalian herbivores. In 1999, we began measuring vegetation structure and p
SGS-LTER Long-Term Monitoring Project: Small Mammals on Trapping Webs on the Central Plains Experimental Range, Nunn, Colorado, USA 1994 -2006, ARS Study Number 118 (Reformatted to a Darwin Core Archive)
This data package is formatted as a Darwin Core Archive (DwC-A, event core). For more information on Darwin Core see https://www.tdwg.org/standards/dwc/. This Level 2 data package was derived from the Level 1 data package found here: https://pasta.lternet.edu/package/metadata/eml/edi/329/2, which was derived from the Level 0 data package found here: https://pasta.lternet.edu/package/metadata/eml/knb-lter-sgs/137/17. The abstract below was extracted from the Level 0 data package and is included for context: This data package was produced by researchers working on the Shortgrass Steppe Long Term Ecological Research (SGS-LTER) Project, administered at Colorado State University. Long-term datasets and background information (proposals, reports, photographs, etc.) on the SGS-LTER project are contained in a comprehensive project collection within the Digital Collections of Colorado (http://digitool.library.colostate.edu/R/?func=collections&collection_id=3429). The data table and associated metadata document, which is generated in Ecological Metadata Language, may be available through other repositories serving the ecological research community and represent components of the larger SGS-LTER project collection. Additional information and referenced materials can be found: http://hdl.handle.net/10217/83452. Small mammals (rabbits, rodents) are integral components of semiarid ecosystems because of their roles as consumers of plants, seeds and arthropods, as soil disturbance agents, and as food for raptors, snakes and mammalian carnivores. Because of their vagility and intermediate trophic position, populations of small mammals may track changes in vegetation and the abiotic environment that may result from shifts in land-use and other anthropogenic disturbances. However, these populations are variable over space and time, and their response to environmental changes may not be immediately apparent given their behavioral flexibility and relatively long life-spans and generatio
SGS-LTER Long-Term Montioring Project: Arthropod Pitfall Trapping on Small Mammal Trapping Webs on the Central Plains Experimental Range, Nunn, Colorado, USA 1998-2006, ARS Study Number 118 (Reformatted to a Darwin Core Archive)
This data package is formatted as a Darwin Core Archive (DwC-A, event core). For more information on Darwin Core see https://www.tdwg.org/standards/dwc/. This Level 2 data package was derived from the Level 1 data package found here: https://pasta.lternet.edu/package/metadata/eml/edi/328/2, which was derived from the Level 0 data package found here: https://pasta.lternet.edu/package/metadata/eml/knb-lter-sgs/134/17. The abstract below was extracted from the Level 0 data package and is included for context: This data package was produced by researchers working on the Shortgrass Steppe Long Term Ecological Research (SGS-LTER) Project, administered at Colorado State University. Long-term datasets and background information (proposals, reports, photographs, etc.) on the SGS-LTER project are contained in a comprehensive project collection within the Digital Collections of Colorado (http://digitool.library.colostate.edu/R/?func=collections&collection_id=3429). The data table and associated metadata document, which is generated in Ecological Metadata Language, may be available through other repositories serving the ecological research community and represent components of the larger SGS-LTER project collection. Additional information and referenced materials can be found: http://hdl.handle.net/10217/83450. With the exception of heteromyids, eg kangaroo rats and pocket mice, most small rodents in shortgrass steppe are omnivorous. Depending on season, arthropods (insects and arachnids) make up 40-85% of the diet of grasshopper mice and thirteen-lined ground squirrels, the most widespread rodents in northern shortgrass steppe. Small mammals are among the most important predators of ground-dwelling macroarthropods and herbivorous insects provide a direct resource link between weather and plant production. Understanding temporal variability in the abundance of arthropods is central to determining the mechanisms that drive small rodent populations. At present, there are no long-
Rodent data from trapping webs in the long-term Small Mammal Exclusion Study (SMES) at Jornada Basin LTER, 1995-2007
This data package contains rodent trapping data from plots with various levels of herbivore exclusion on the Jornada Experimental Range (JER) and Chihuahuan Desert Rangeland Research Center (CDRRC) lands. Study sites were established in 1995; one in black grama grassland and the other in creosotebush shrubland to compare the impact of herbivores on ecosystem processes between these vegetation types. Parallel studies were established at the Sevilleta LTER site (New Mexico, USA) and Mapimi Biosphere Reserve (Durango, Mexico). Each study site is 1 km by 0.5 km in area. Three replicate rodent trapping webs and four replicate experimental blocks were randomly located at each study site. Rodent trapping webs were used to measure rodent population density and species diversity over time, while the experimental blocks measure vegetation responses to herbivore exclusion treatments including a) all mammalian herbivores, including cattle, lagomorphs, and rodents, b) lagomorphs and cattle only, c) cattle only, and d) control accessible to all herbivores. Rodent populations were sampled from each of the three webs at each study site during overnight trapping campaigns twice per year, in the early (April-May) and late (September-October) summer between 1995 and 2007 (trapping study terminated after October 2007). During each trapping campaign, live-traps were left open for three consecutive nights, and captured animals were recorded on the three subsequent mornings. Each animal caught was identified, measured, and released at the same location where it was captured. This study is complete.
Supplementary Files for RIDE Review of Juxta Web Service, LERA, and Variance Viewer
<p>Test datasets and result files for review on web based collation tools <a href="http://www.juxtacommons.org/"><em>Juxta Web Service</em></a>, <a href="https://lera.uzi.uni-halle.de/"><em>LERA</em></a>, and <a href="http://variance-viewer.informatik.uni-wuerzburg.de/Variance-Viewer/"><em>Variance Viewer</em></a> (see <a href="https://ride.i-d-e.de">RIDE</a> issue on Tools and Environments for Digital Scholarly Editing).</p> <p>The first dataset (lorem-*) are constructed dummy texts in TEI, based on filler texts. They cover some basic cases of textual variance.</p> <p>The second dataset (hamlet-*) is based on a TEI encoded version of Shakespeare's <em>Hamlet</em>, taken from <a href="http://quartos.org">The Shakespeare Quartos Archive</a> (CC BY-NC 2.0), with a simplified baseline encoding.</p> <p>The result files (result-*) were produced with the web based collation tools mentioned above. The LERA results are available in PDF, while the other two tools provided TEI-XML. There is an extra configuration file for Variance Viewer.</p>
Magento eCommerce Web Development Services in India
<p><a href="http://www.magentoindia.com/"><strong>Magento India</strong></a> - is an independent <a href="http://www.magentoindia.com/magento-development.html"><strong>Magento eCommerce development service provider in India</strong></a>. One can avail here one-stop solutions for Magento development, such as custom Magento theme development, custom Magento web development, our Magento eCommerce development services at a reliable price. Counting with 10 Years+ expertise in web development, our team completed successfully hundreds of projects with customer requirement and satisfaction. At the same time, one can also hire Magento developers in India to get the task done within the time frame and yet at very reasonable hourly rates or fixed price too. So, if you are looking for Magento development agencies then you are at the right place to approach us to deliver the best possible services and customer support with your comfort.</p> <p>For consultation contact: 01204380622, +91-9650657773, +91-9818460005.</p> <p> </p>
Open Source Development Services Company - Global Web Seller
<p><a href="http://globalwebseller.com/open-source-customization.html"><strong>Open Source Development</strong></a> - is very demandable from the customer point of view. Open source development consists of various web development technologies and very user-friendly and also the best option for a client with respect to its customization and cost too. So, I would like to introduce a leading and specialized <a href="http://globalwebseller.com/"><strong>Web development company</strong></a> Global Web Seller an ISO 9001:2008 certified for prior web development service for open source. </p> <p>They provide the following open source development services - </p> <p>1. Joomla</p> <p>2. PHP</p> <p>3. Wordpress</p> <p>4. OsCommerce</p> <p>5. Drupal</p> <p>6. Magento</p> <p>7. Laravel</p> <p>8. Java</p> <p>9. Zend Framework etc.</p> <p>So, here you can avail of complete web designing and development with the best pricing option. All their web developers are highly experienced and capable to provide an innovative and custom eCommerce website, matrimonial portal, real estate portal development, a shopping website and a lot more. </p> <p>Get in touch - </p> <p>Contact Information:</p> <p>01204380622, +91-9818460005, +91-9650657773</p> <p>Write email to your query - sales@globalwebseller.com</p> <p> </p> <p> </p> <p> </p> <p> </p>
WICE-DB - Web Instrument Count Experiment
<p>This dataset consist of 12 hand selected musical stimuli to evaluate source count experiments.</p> <p>However, to better study the influence of vibrato we require extended control over certain parameters such as note duration, vibrato duration, exact fundamental frequency, vibrato rate, vibrato extend, reproducibility, loudness or expression.</p> <p>As mentioned in the previous chapter, vibrato techniques vary across instruments. Instruments such as violin and saxophone are known for their distinct frequency modulations. Other instruments such as the English horn and the flute are more close to amplitude modulations.</p> <p>We generated the notes using a software sampler which allows us to control the parameters such as the vibrato. All our test stimuli have a duration of three seconds. Items were equalized in loudness by using an iterative calculation of the loudness algorithm of the time-varying Zwicker model. We rendered 29 notes of C4, resulting in 841 unique unison instrument mixtures per pitch class.</p>
The MedAfriCarbon radiocarbon database and web application. Archaeological dynamics in Mediterranean Africa, ca. 9600-700 BC
<p>MedAfriCarbon radiocarbon database and web app are outcomes of the <em>MedAfrica Project —Archaeological deep history and dynamics of Mediterranean Africa, ca. 9600-700 BC</em>. The dataset presented here includes a collection of <strong>1584</strong> calibrated archaeological 14C dates from <strong>1587</strong> samples collected from <strong>368</strong> sites located in Mediterranean Africa (plus some additional dates whose published information is incomplete). The majority of the dates are linked to cultural and environmental variables, notably the presence/absence of different domestic/wild species and specific material culture.</p> <ul> <li><strong>1.0.3</strong> _27 Feb 2020_– Official public JOAD release (includes JOAD DOI and volume numbers). Includes a few minor error corrections.</li> <li><strong>1.0</strong> _30 Jan 2020_— First public release of the dataset on Zenodo.</li> </ul>
Data Science in Biomedicine - Web of Science datasets
<p>Datasets from the Web of Science search for the number of publications associated with the topics "Data Science", "Big Data" and "Cloud Computing" from 2004 to 2019 in 9 different countries.</p>
Top: puffs of cornstarch reveal dense and varied tiny cryptic webs in the Gaoligongshan. Shown here upper left to lower right are a symphytognathid Patu jidanweishi sp. n., a mysmenid Gaoligonga changya gen. n., sp. n., and an unidentified linyphiid. Bottom: this misty mountain landscape at QiQi is typical of the Gaoligongshan in The symphytognathoid spiders of the Gaoligongshan, Yunnan, China (Araneae: Araneoidea): Systematics and diversity of micro-orbweavers
Top: puffs of cornstarch reveal dense and varied tiny cryptic webs in the Gaoligongshan. Shown here upper left to lower right are a symphytognathid Patu jidanweishi sp. n., a mysmenid Gaoligonga changya gen. n., sp. n., and an unidentified linyphiid. Bottom: this misty mountain landscape at QiQi is typical of the Gaoligongshan
Structural Profiling of Web Sites in the Wild
<p>The dataset contains and processes results of a large-scale survey of 708 websites, made in December 2019, in order to measure various features related to their size and structure: DOM tree size, maximum degree, depth, diversity of element types and CSS classes, among others. The goal of this research is to serve as a reference point for studies that include an empirical evaluation on samples of web pages.</p> <p>See the Readme.md file inside the archive for more details about its contents.</p>
Dataset for: The Evolution of the Manosphere Across the Web
<p><strong>The Evolution of the Manosphere Across the Web</strong></p> <p>We make available data related to subreddit and standalone forums from the manosphere.</p> <p>We also make available Perspective API annotations for all posts.</p> <p>You can find the code in <a href="https://github.com/manoelhortaribeiro/manosphere_analysis">GitHub</a>.</p> <p>Please cite this paper if you use this data:</p> <pre><code>@article{ribeiroevolution2021, title={The Evolution of the Manosphere Across the Web}, author={Ribeiro, Manoel Horta and Blackburn, Jeremy and Bradlyn, Barry and De Cristofaro, Emiliano and Stringhini, Gianluca and Long, Summer and Greenberg, Stephanie and Zannettou, Savvas}, booktitle = {{Proceedings of the 15th International AAAI Conference on Weblogs and Social Media (ICWSM'21)}}, year={2021} } </code></pre> <p><strong>1. Reddit data</strong></p> <p>We make available data for forums and for relevant subreddits (56 of them, as described in <code>subreddit_descriptions.csv</code>).<br> These are available, 1 line per post in each subreddit Reddit in <code>/ndjson/reddit.ndjson</code>.<br> A sample for example is:</p> <pre><code>{ "author": "Handheld_Gaming", "date_post": 1546300852, "id_post": "abcusl", "number_post": 9.0, "subreddit": "Braincels", "text_post": "Its been 2019 for almost 1 hour And I am at a party with 120 people, half of them being foids. The last year had been the best in my life. I actually was happy living hope because I was redpilled to the death. \n\nNow that I am blackpilled I see that I am the shortest of all men and that I am the only one with a recessed jaw. \n\nIts over. Its only thanks to my age old friendship with chads and my social skills I had developed in the past year that a lot of men like me a lot as a friend.\n\nNo leg lengthening syrgery is gonna save me. Ignorance was a bliss. Its just horror now seeing that everyone can make out wirth some slin hoe at the party. \n\n I actually feel so unbelivably bad for turbomanlets. Life as an unattractive manlet is a pain, I cant imagine the hell being an ugly turbomanlet is like. I would have roped instsntly if I were one. Its so unfair. \n\nTallcels are fakecels and they all can (and should) suck my cock. \n\nIf I were 17cm taller my life would be a heaven and I would be the happiest man alive. \n\nJust cope and wait for affordable body tranpslants.", "thread": "t3_abcusl" } </code></pre> <p><strong>2. Forums</strong></p> <p>We here describe the <code>.sqlite</code> and <code>.ndjson</code> files that contain the data from the following forums.</p> <pre><code>(avfm) --- https://d2ec906f9aea-003845.vbulletin.net (incels) --- https://incels.co/ (love_shy) --- http://love-shy.com/lsbb/ (redpilltalk) --- https://redpilltalk.com/ (mgtow) --- https://www.mgtow.com/forums/ (rooshv) --- https://www.rooshvforum.com/ (pua_forum) --- https://www.pick-up-artist-forum.com/ (the_attraction) --- http://www.theattractionforums.com/ </code></pre> <p>The files are in folders <code>/sqlite/</code> and <code>/ndjson</code>.</p> <p><strong><code>2.1 .sqlite</code></strong></p> <p>All the tables in the <code>sqlite.</code> datasets follow a very simple <code>{key:value}</code> format.<br> Each <code>key</code> is a thread name (for example <code>/threads/housewife-is-like-a-job.123835/</code>) and each value is a python dictionary or a list.<br> This file contains three tables:</p> <ul> <li> <p><code>idx</code> each key is the relative address to a thread and maps to a post. Each post is represented by a dict:</p> <pre><code>"type": (list) in some forums you can add a descriptor such as [RageFuel] to each topic, and you may also have special types of posts, like sticked/pool/locked posts. "title": (str) title of the thread; "link": (str) link to the thread; "author_topic": (str) username that created the thread; "replies": (int) number of replies, may differ from number of posts due to difference in crawling date; "views": (int) number of views; "subforum": (str) name of the subforum; "collected": (bool) indicates if raw posts have been collected; "crawled_idx_at": (str) datetime of the collection. </code></pre> </li> <li> <p><code>processed_posts</code> each key is the relative address to a thread and maps to a list with posts (in order). Each post is represented by a dict:</p> <pre><code>"author": (str) author's username; "resume_author": (str) author's little description; "joined_author": (str) date author joined; "messages_author": (int) number of messages the author has; "text_post": (str) text of the main post; "number_post": (int) number of the post in the thread; "id_post": (str) unique post identifier (depends), for sure unique within thread; "id_post_interaction": (list) list with other posts ids this post quoted; "date_post": (str) datetime of the post, "links": (tuple) nice tuple with the url parsed, e.g. ('https', 'www.youtube.com', '/S5t6K9iwcdw'); "thread": (str) same as key; "crawled_at": (str) datetime of the collection. </code></pre> </li> <li> <p><code>raw_posts</code> each key is the relative address to a thread and maps to a list with unprocessed posts (in order). Each post is represented by a dict:</p> <pre><code>"post_raw": (binary) raw html binary; "crawled_at": (str) datetime of the collection. </code></pre> </li> </ul> <p><strong><code>2.2 .ndjson</code></strong></p> <p>Each line consists of a json object representing a different comment with the following fields:</p> <pre><code> "author": (str) author's username; "resume_author": (str) author's little description; "joined_author": (str) date author joined; "messages_author": (int) number of messages the author has; "text_post": (str) text of the main post; "number_post": (int) number of the post in the thread; "id_post": (str) unique post identifier (depends), for sure unique within thread; "id_post_interaction": (list) list with other posts ids this post quoted; "date_post": (str) datetime of the post, "links": (tuple) nice tuple with the url parsed, e.g. ('https', 'www.youtube.com', '/S5t6K9iwcdw'); "thread": (str) same as key; "crawled_at": (str) datetime of the collection. </code></pre> <p><strong>3. Perspective</strong></p> <p>We also run each post and reddit post through perspective, the files are located in the <code>/perspective/</code> folder.<br> They are compressed with gzip.<br> One example output</p> <pre><code>{ "id_post": 5200, "hate_output": { "text": "I still can\u2019t wrap my mind around both of those articles about these c\~\~\~s sleeping with poor Haitian Men. Where\u2019s the uproar?, where the hell is the outcry?, the \u201cpig\u201d comments or the \u201ccreeper comments\u201d. F\~\~\~ing hell, if roles were reversed and it was an article about Men going to Europe where under 18 sex in legal, you better believe they would crucify the writer of that article and DEMAND an apology by the paper that wrote it.. This is exactly what I try and explain to people about the double standards within our modern society. A bunch of older women, wanna get their kicks off by sleeping with poor Men, just before they either hit or are at menopause age. F~~~ing unreal, I\u2019ll never forget going to Sweden and Norway a few years ago with one of my buddies and his girlfriend who was from there, the legal age of consent in Norway is 16 and in Sweden it\u2019s 15. I couldn\u2019t believe it, but my friend told me \u201c hey, it\u2019s normal here\u201d . Not only that but the age wasn\u2019t a big different in other European countries as well. One thing i learned very quickly was how very Misandric Sweden as well as Denmark were.", "TOXICITY": 0.6079781, "SEVERE_TOXICITY": 0.53744453, "INFLAMMATORY": 0.7279288, "PROFANITY": 0.58842486, "INSULT": 0.5511079, "OBSCENE": 0.9830818, "SPAM": 0.17009115 } } </code></pre> <p><strong>4. Working with sqlite</strong></p> <p>A nice way to read some of the files of the dataset is using <code>SqliteDict</code>, for example:</p> <pre><code>from sqlitedict import SqliteDict processed_posts = SqliteDict("./data/forums/incels.sqlite", tablename="processed_posts") for key, posts in processed_posts.items(): for post in posts: # here you could do something with each post in the dataset pass </code></pre> <p><strong>5. Helpers</strong></p> <p>Additionally, we provide two <code>.sqlite</code> files that are helpers used in the analyses.<br> These are related to reddit, and not to the forums!<br> They are:</p> <ul> <li> <p><code>channel_dict.sqlite</code> a sqlite where each key corresponds to a subreddit and values are lists of dictionaries users who posted on it, along with timestamps.</p> </li> <li> <p><code>author_dict.sqlite</code> a sqlite where each key corresponds to an author and values are lists of dictionaries of the subreddits they posted on, along with timestamps.</p> </li> </ul> <p>These are used in the paper for the migration analyses.</p> <p><strong>6. Examples and particularities for forums</strong></p> <p>Although we did our best to clean the data and be consistent across forums, this is not always possible. In the following subsections we talk about the particularities of each forum, directions to improve the parsing which were not pursued as well as give some examples on how things work in each forum.</p> <p><strong>6.1 incels</strong></p> <p>Check out an archived version of the <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/incels_front_page.png">front page</a>, the <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/incels_thread_page.png">thread page</a> and a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/incels_post_page.png">post page</a>, as well as a dump of the data stored for a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/examples/incels_thread_page_example.txt">thread page</a> and a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/examples/incels_post_page_example.txt">post page</a>.</p> <ul> <li> <p><em>types</em>: for the incel forums the special types associated with each thread in the <code>idx</code> table are “Sticky”, “Pool”, “Closed”, and the custom types added by users, such as <code>[LifeFuel]</code>. These last ones are all in brackets. You can see some examples of these in the on the example thread page.</p> </li> <li> <p><em>quotes</em>: quotes in this forum were quite nice and thus, all quotations are deterministic.</p> </li> </ul> <p><strong>6.2 LoveShy</strong></p> <p>Check out an archived version of the <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/love_shy_front_page.png">front page</a>, the <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/love_shy_thread_page.png">thread page</a> and a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/love_shy_thread_page.png">post page</a>, as well as a dump of the data stored for a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/examples/love_shy_thread_page_example.txt">thread page</a> and a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/examples/love_shy_post_page_example.txt">post page</a>.</p> <ul> <li> <p><em>types</em>: no types were parsed. There are some rules in the forum, but not significant.</p> </li> <li> <p><em>quotes</em>: quotes were obtained from exact text+author match, or author match + a jaccard similarity of more than 0.9 in the tokens of the text split on spaces.</p> </li> <li> <p><em>other</em>: this forum apparently deleted a bunch of posts (or hid them from new users).</p> </li> </ul> <p><strong>6.3 MGTOW</strong></p> <p>Check out an archived version of the <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/mgtow_front_page.png">front page</a>, the <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/mgtow_thread_page.png">thread page</a> and a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/mgtow_thread_page.png">post page</a>, as well as a dump of the data stored for a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/examples/mgtow_thread_page_example.txt">thread page</a> and a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/examples/mgtow_post_page_example.txt">post page</a>.</p> <ul> <li> <p><em>types</em>: no types were parsed. There are some rules in the forum, but not significant.</p> </li> <li> <p><em>quotes</em>: quotes in this forum were quite nice and thus, all quotations are deterministic.</p> </li> <li> <p><em>other</em>: collecting users for this forum was hard because info like number of threads and join date do not show on the threads. Thus there was a 2-step collection. Some users may be missing because of this.</p> </li> </ul> <p><strong>6.4 A Voice for Men</strong></p> <p>Check out an archived version of the <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/avfm_front_page.png">front page</a>, the <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/avfm_thread_page.png">thread page</a> and a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/avfm_thread_page.png">post page</a>, as well as a dump of the data stored for a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/examples/avfm_thread_page_example.txt">thread page</a> and a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/examples/avfm_post_page_example.txt">post page</a>.</p> <ul> <li> <p><em>types</em>: “sticky” is the only existing type.</p> </li> <li> <p><em>quotes</em>: quotes in this forum were quite nice and thus, all quotations are deterministic.</p> </li> </ul> <p><strong>6.5 Red Pill Talk (formerly <a href="http://sluthate.com">sluthate.com</a>)</strong></p> <p>Check out an archived version of the <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/avfm_front_page.png">front page</a>, the <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/avfm_thread_page.png">thread page</a> and a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/avfm_thread_page.png">post page</a>, as well as a dump of the data stored for a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/examples/avfm_thread_page_example.txt">thread page</a> and a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/examples/avfm_post_page_example.txt">post page</a>.</p> <ul> <li> <p><em>types</em>: no types were parsed.</p> </li> <li> <p><em>quotes</em>: quotes were obtained from exact text+author match, or author match + a jaccard similarity of more than 0.9 in the tokens of the text split on spaces.</p> </li> </ul> <p><strong>6.6 The Attraction</strong></p> <p>Check out an archived version of the <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/the_attraction_front_page.png">front page</a>, the <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/the_attraction_thread_page.png">thread page</a> and a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/the_attraction_thread_page.png">post page</a>, as well as a dump of the data stored for a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/examples/the_attraction_thread_page_example.txt">thread page</a> and a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/examples/the_attraction_post_page_example.txt">post page</a>.</p> <ul> <li> <p><em>types</em>: “sticky” is the only existing type.</p> </li> <li> <p><em>quotes</em>: quotes in this forum were quite nice and thus, all quotations are deterministic.</p> </li> <li> <p><em>other</em>: there is the possibility to parse stuff like age and location, but this was not done.</p> </li> </ul> <p><strong>6.7 PUA Forum</strong></p> <p>Check out an archived version of the <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/pua_forum_front_page.png">front page</a>, the <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/pua_forum_thread_page.png">thread page</a> and a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/pua_forum_thread_page.png">post page</a>, as well as a dump of the data stored for a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/examples/pua_forum_thread_page_example.txt">thread page</a> and a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/examples/pua_forum_post_page_example.txt">post page</a>.</p> <ul> <li> <p><em>types</em>: There are several types parsed: Sticky, Announcement, General Announcement, Locked, to name a few.</p> </li> <li> <p><em>quotes</em>: quotes were obtained from exact text match, or a jaccard similarity of more than 0.9 in the tokens of the text split on spaces. There are actually some problems with quotes with other quotes inside (check example .txt).</p> </li> </ul> <p><strong>6.8 Rooshv</strong></p> <p>Check out an archived version of the <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/rooshv_front_page.png">front page</a>, the <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/rooshv_thread_page.png">thread page</a> and a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/screenshots/pua_forum_thread_page.png">post page</a>, as well as a dump of the data stored for a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/examples/rooshv_thread_page_example.txt">thread page</a> and a <a href="https://github.com/manoelhortaribeiro/manosphere_analysis/blob/master/data/forums/examples/rooshv_post_page_example.txt">post page</a>.</p> <ul> <li> <p><em>types</em>: Two types parsed: <code>sticky</code>, <code>announcement</code>.</p> </li> <li> <p><em>quotes</em>: quotes in this forum were quite nice and thus, all quotations are deterministic.</p> </li> <li> <p><em>other</em>: lots of posts have some problem with dates, which is quite annoying (~5%). Some threads could not be captured also, website behaves a bit weird. Some don’t exist.</p> </li> </ul>
Data from: Impacts of nutrient subsidies on salt marsh arthropod food webs: a latitudinal survey
Anthropogenic nutrient inputs into native ecosystems cause fluctuations in resources that normally limit plant growth, which has important consequences for associated foodwebs. Such inputs from agricultural and urban habitats into nearby natural systems are increasing globally and can be highly variable. Despite the global increase in anthropogenically-derived nutrient inputs into native ecosystems, the consequences of variation in subsidy amount on native plants and their associated foodwebs are poorly known. Salt marshes represent an ideal system to address the differential impacts of nutrient inputs on ecosystem and community dynamics because human development and other anthropogenic activities lead to recurrent introductions of nutrients into these natural systems. Previously, we have found in manipulative experiments that arthropod abundance increases in response to nutrient enrichment, with predators being the trophic group most strongly affected. We conducted a survey of Atlantic coastal Spartina marshes to test whether such local responses are indicative of responses at a landscape level. We examined the most abundant arthropod species associated with Spartina coastal marshes that receive variable amounts of anthropogenic nitrogen, and tested how this response varied across different arthropod functional groups (herbivores, epigeic feeders, and predators). Similar to what we found at a local scale, nutrient subsidies alter the trophic structure of the arthropod assemblage by changing the relative abundances of various feeding groups. Variable responses among predators to nitrogen density could be partly explained by diet breadth (e.g. generalists vs. specialists). Herbivores had a negative response to increasing plant nitrogen density; specialist predators tracked their herbivore prey and thus also responded negatively to nitrogen density. However, generalists were not negatively affected by nitrogen density and indeed some generalist predators responded positively to nitrogen density. Thus, the overall predator-to-herbivore ratio was also positively associated with nitrogen density. Our research helps us to understand how long-term nutrient enrichment of native ecosystems by human activities affects arthropod assemblages and foodweb dynamics.
Utilização do Scrum no desenvolvimento de uma aplicação web: um Estudo de Caso
<p>Vídeo de apresentação do trabalho entitulado "Utilização do Scrum no desenvolvimento de uma aplicação web: um Estudo de Caso", a ser apresentado no Fórum de Graduação na IV Escola Regional de Engenharia de Software (ERES 2020) - Sessão Técnica 6.</p>
Disentangling ecological and taphonomic signals in ancient food webs
<p>Analyses of ancient food webs reveal important paleoecological processes and responses to a range of perturbations throughout Earth's history, such as climate change. These responses can inform our forecasts of future biotic responses to similar perturbations. However, previous analyses of ancient food webs rarely accounted for key differences between modern and ancient community data, particularly selective loss of soft-bodied taxa during fossilization. To consider how fossilization impacts inferences of ancient community structure we (1) analyzed node-level attributes to identify correlations between ecological roles and fossilization potential and (2) applied selective information loss procedures to food web data for extant systems. We found that selective loss of soft-bodied organisms has predictable effects on the trophic structure of "artificially fossilized" food webs, because these organisms occupy unique, consistent food web positions. Fossilized food webs misleadingly appear less stable (i.e., more prone to trophic cascades), with less predation and an overrepresentation of generalist consumers. We also found that ecological differences between soft- and hard-bodied taxa—indicated by distinct positions in modern food webs—are recorded in an Early Eocene web, but not in Cambrian webs. This suggests that ecological differences between the groups have existed for ≥ 48 million years. Our results indicate that accounting for soft-bodied taxa is vital for accurate depictions of ancient food webs. However, the consistency of information loss trends across the analyzed food webs means it is possible to predict how the selective loss of soft-bodied taxa affects food web metrics, which can permit better modeling of ancient communities.</p>
Automated prediction of visual complexity of web pages: Tools and evaluations
<p>This dataset includes the screenshots of the web pages used for the evaluation of ViCRAM which is described in the following paper:</p> <p>Eleni Michailidou, Sukru Eraslan, Yeliz Yesilada, and Simon Harper. 2020. Automated Prediction of Visual Complexity of Web Pages: Tools and Evaluations. International Journal of Human-Computer Studies (SCI-E, SSCI), 145, 102523.</p>
Data set of the article: Using Machine Learning for Web Page Classification in Search Engine Optimization
<p>Data of investigation published in the article: "Using Machine Learning for Web Page Classification in Search Engine Optimization"</p> <p>Abstract of the article:</p> <p>This paper presents a novel approach of using machine learning algorithms based on experts’ knowledge to classify web pages into three predefined classes according to the degree of content adjustment to the search engine optimization (SEO) recommendations. In this study, classifiers were built and trained to classify an unknown sample (web page) into one of the three predefined classes and to identify important factors that affect the degree of page adjustment. The data in the training set are manually labeled by domain experts. The experimental results show that machine learning can be used for predicting the degree of adjustment of web pages to the SEO recommendations—classifier accuracy ranges from 54.59% to 69.67%, which is higher than the baseline accuracy of classification of samples in the majority class (48.83%). Practical significance of the proposed approach is in providing the core for building software agents and expert systems to automatically detect web pages, or parts of web pages, that need improvement to comply with the SEO guidelines and, therefore, potentially gain higher rankings by search engines. Also, the results of this study contribute to the field of detecting optimal values of ranking factors that search engines use to rank web pages. Experiments in this paper suggest that important factors to be taken into consideration when preparing a web page are page title, meta description, H1 tag (heading), and body text—which is aligned with the findings of previous research. Another result of this research is a new data set of manually labeled web pages that can be used in further research. </p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.