Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
5,805
datasets available to search
ShareScore release 0.9.0
Dataset results
5,805 results for “Data model”
CLPX-Model: Land Data Assimilation System (LDAS) Data, Version 1
The LDAS data set contains 43 model and observation-based fields produced by the LDAS uncoupled modeling system at the NASA Goddard Space Flight Center using the Mosaic Land Surface Model (LSM).
Daily 4 km Gridded SWE and Snow Depth from Assimilated In-Situ and Modeled Data over the Conterminous US, Version 1
This data set provides daily 4 km snow water equivalent (SWE) and snow depth over the conterminous United States. It was developed at the University of Arizona (UA) under the support of the NASA MAP and SMAP Programs. The data were created by assimilating in-situ snow measurements from the National Resources Conservation Service's SNOTEL network and the National Weather Service's COOP network with modeled, gridded temperature and precipitation data from PRISM.
LBA-ECO LC-15 SRTM30 Digital Elevation Model Data, Amazon Basin: 2000
This dataset provides a subset of the SRTM30 Digital Elevation Model (DEM) elevation and standard deviation data for the Amazon Basin. SRTM30 is a near-global digital elevation model (DEM) comprising a combination of data from the Shuttle Radar Topography Mission (SRTM), flown in February, 2000, and the earlier U.S. Geological Survey's GTOPO30 data set. The SRTM30 resolution is 30 arc-sec or about 1 km. In processing the SRTM data, to combine with GTOPO30, the data were resampled from 3 arc-sec to 30 arc-sec. Provided here are the mean elevation and the standard deviation (STD) of the data points used in the averaging. The STD is thus an indication of topographic roughness useful in some applications.
NACP MsTMIP: Global and North American Driver Data for Multi-Model Intercomparison
This data set provides environmental data that have been standardized and aggregated for use as input to carbon cycle models at global (0.5-degree resolution) and regional (North America at 0.25-degree resolution) scales. The data were compiled from selected sources (Table 2) and integrated into gridded global and regional collections of climatology variables (precipitation, air temperature, air specific humidity, air relative humidity (NA only), pressure, downward longwave radiation, downward shortwave radiation, and wind speed), time-varying atmospheric CO2 concentrations, time-varying nitrogen deposition, biome fraction and type, land-use and land-cover change, C3/C4 grasses fractions, major crop distribution, phenology, multiple soil characteristics, and a land-water mask. The temporal ranges of the data are sufficient for carbon cycle model simulations from 1801 to 2010. These data were compiled specifically for the North American Carbon Program (NACP) Multi-Scale Synthesis and Terrestrial Model Intercomparison Project (MsTMIP) as the prescribed model input driver data (Huntzinger et al., 2013). The driver data were used by 22 terrestrial biosphere models to run baseline and sensitivity simulations. The standardized data provided consistent model inputs to minimize the inter-model variability caused by differences in environmental drivers and initial conditions. Together with the sensitivity simulations, the standardized input data enable better interpretation and quantification of structural and parameter uncertainties of model estimates. Data are provided in Climate and Forecast (CF) metadata convention compliant (version 1.4) netCDF-4 file formats. There are 3,152 *.nc4 data files with this data set.
ABoVE: Light-Curve Modelling of Gridded GPP Using MODIS MAIAC and Flux Tower Data
This dataset contains gridded estimations of daily ecosystem Gross Primary Production (GPP) in grams of carbon per day at a 1 km2 spatial resolution over Alaska and Canada from 2000-01-01 to 2018-01-01. Daily estimates of GPP were derived from a light-curve model that was fitted and validated over a network of ABoVE domain Ameriflux flux towers then upscaled using MODIS Multi-Angle Implementation of Atmospheric Correction (MAIAC) data to span the extended ABoVE domain. In general, the methods involved three steps; the first step involved collecting and processing mainly carbon-flux site-level data, the second step involved the analysis and correction of site-level MAIAC data, and the final step developed a framework to produce large-scale estimates of GPP. The light-curve parameter model was generated by upscaling from flux tower sub-daily temporal resolution by deconvolving the GPP variable into 3 components: the absorbed photosynthetically active radiation (aPAR), the maximum GPP or maximum photosynthetic capacity (GPPmax), and the photosynthetic limitation or amount of light needed to reach maximum capacity (PPFDmax). GPPmax and PPFDmax were related to satellite reflectance measurements sampled at the daily scale. GPP over the extended ABoVE domain was estimated at a daily resolution from the light-curve parameter model using MODIS MAIAC daily reflectance as input. This framework allows large-scale estimates of phenology and evaluation of ecosystem sensitivity to climate change.
NACP Site: Terrestrial Biosphere Model and Aggregated Flux Data in Standard Format
This data set provides standardized output variables for gross primary productivity (GPP), net ecosystem exchange (NEE), leaf area index (LAI), ecosystem respiration (Re), latent heat flux (LE), and sensible heat flux (H) from 24 terrestrial biosphere models for 47 eddy covariance flux tower sites in North America. Each model used standardized input data for each flux tower site (i.e., gap-filled, locally observed weather; land use history; and other site specific data) and followed standard model setup and spinup procedures. The files also contain gap-filled observations and total uncertainty estimates. The data set was compiled for the North American Carbon Program (NACP) Site-Level Synthesis for use in model inter-comparison and assessment of how well the models simulate carbon processes across vegetation types and environmental conditions in North America. There is one compressed (.zip) file with this data set. When expanded, the .zip file contains model output data for one variable at one site. The model output and observations are available at the native half-hourly time step, or in daily, monthly, and annual aggregations, in comma-separated text (.csv) format.
NACP: Climate Data Inputs (3-hourly) for Community Land Model, Western USA, 1979-2015
This dataset provides sub-daily, high-resolution, climate data inputs including temperature, precipitation, near surface specific humidity, incoming short-wave radiation, and near-surface wind speed over 11 states of the western USA. States included are Arizona, California, Colorado, Idaho, Montana, Nevada, New Mexico, Oregon, Utah, Washington, and Wyoming. These data were derived for use in the Community Land Model (CLM v4.5) and are at 3-hourly temporal and 4 x 4 km spatial resolutions for the 1979 through 2015 time period. The source for observational data was METDATA (now called GRIDMET), at a daily resolution. Modeling efforts using these data estimated annual carbon stocks, fluxes, and productivity across the western United States.
NACP Site: Terrestrial Biosphere Model Output Data in Original Format
This data set contains the original model output data submissions from the 24 terrestrial biosphere models (TBM) that participated in the North American Carbon Program (NACP) Site-Level Synthesis. The model teams generated estimates for, but not limited to, a minimum of six variables, including gross primary productivity (GPP), net ecosystem exchange (NEE), leaf area index (LAI), ecosystem respiration (Re), latent heat flux (LE), and sensible heat flux (H) for each of 47 selected eddy covariance flux tower sites across North America. Participating modeling teams followed the NACP Site Synthesis Protocol (site_synthesis_protocol_v7.pdf), which covers procedures, plans, and infrastructure for the site-level analyses. File format and units conversions of several data submissions were made by the MAST-DC to produce NetCDF files of consistent content and structure for all 24 TBM outputs. The model outputs are structured as described in Appendix A: Model Output Variables, of the Site Synthesis Protocol. In addition, MAST-DC processed these original model submissions to derive uniquely processed and formatted data files for model inter-comparison and evaluation (NACP Site: Terrestrial Biosphere Model and Aggregated Flux Data in Standard Format). This related data set provides GPP, NEE, LAI, Re, LE, and sensible heat (H) model output variables at the native half-hourly time step, and in daily, monthly, and annual aggregations. The related data set also contains gap-filled observations and total uncertainty estimates at the same time steps.There are 24 compressed (*.zip) files with this data set -- one file for each model. When expanded, the .zip files contain model output data files for flux tower sites in NetCDF and some in text formats.
LBA-ECO CD-32 LBA Model Intercomparison Project (LBA-MIP) Forcing Data
The source meteorological observations for the forcing data, from the nine Brazilian flux towers, were recently published as Saleska, et al. (2013). See related data sets. These source data were gap-filled according to the LBA-MIP standard protocol. Note that the CAX forest tower was not included in the MIP. See the companion file driver_data.pdf for additional gap-filling information.There are 34 data products with this data set and they are provided in both text (.txt) and ALMA-compliant NetCDF (.nc) formats. The files have been compressed into nine *.zip files according to site.
NASA Ocean Biogeochemical Model assimilating satellite chlorophyll data global daily VR2017 (NOBM_DAY) at GES DISC
This is the assimilated daily data from NASA Ocean Biogeochemical Model (NOBM). The NOBM is a comprehensive, interactive ocean biogeochemical model coupled with a circulation and radiative model in the global oceans (Gregg and Casey, 2007). It spans the domain from -84 to 72 degree latitude in increments of 1.25 degree longitude by 2/3 degree latitude, including only open ocean areas where bottom depth > 200m. NOBM contains 4 phytoplankton groups, 4 nutrient groups, a single herbivore group, and 3 detrital pools, and the major ocean carbon components, dissolved organic and inorganic carbon (DOC and DIC).
GEOS-5 FP-IT 3D Time-Averaged Model-Layer Assimilated Data Geo-Colocated to OMI/Aura UV2 1-Orbit L2 Swath 13x24km V4 (OMUFPMET) at GES DISC
The GEOS-5 FP-IT 3D Time-Averaged Model-Layer Assimilated Data Geo-Colocated to OMI/Aura UV2 1-Orbit L2 Swath 13x24km (OMUFPMET) product provides selected meteorlogical fields from the GEOS-5 Forward Processing for Instrument Teams (FP-IT) assimilated product produced by the Global Modeling and Assimilation Office (GMAO) co-located in space and time with the OMI UV-2 swath.The fields in this product include layer pressure thickness, surface pressure, vertical temperature profiles, surface potential, and mid-layer pressure along with geolocation info. The OMI team also provides a corresponding product for the OMI VIS swath, OMVFPMET. The OMI ancillary products were developed to provide supplementary information for use with the OMI collection 4 L1B data sets. The original GEOS-5 FP-IT data are reported on a 0.625 deg longitude by 0.5 deg latitude grid, whereas the OMI UV-2 spatial resolution is 13km x 24km at nadir.The OMUFPMET files are in netCDF4 format which is compatible with most netCDF and HDF5 readers and tools. Each file is approximately 45mb in size. The lead for this product is Zachary Fasnacht of SSAI. Joanna Joiner is the responsible NASA official.
ATom: Simulated Data Stream for Modeling ATom-like Measurements
This dataset provides a simulated data stream representative of an Atmospheric Tomography mission (ATom) data collection flight and also modeled reactivities for ozone (O3) production and loss and methane (CH4) loss from six global atmospheric chemistry models: CAM, GEOS-Chem, GFDL, GISS-E2.1, GMI, and UCI. The simulated data include concentrations of selected atmospheric trace gases for 14,880 air parcels along a simulated north-south ATom flight path along 180-degrees longitude over the Pacific basin. Each of the six models produced ozone production and loss and methane loss reactivities initialized using the simulated data beginning with five different days in August (8-01, 8-06, 8-11, 8-16, 8-21). Modeled years for each individual model varied from 1997 to 2016.
NARSTO 1998 Model-Intercomparison Study Verification Data: NARSTO-Northeast 1995 Surface Ozone, NO, and NOx Langley Data Center Data Set
NARSTO_NE_MODEL is the North American Research Strategy for Tropospheric Ozone (NARSTO) 1998 Model-Intercomparison Study Verification Data: NARSTO-Northeast 1995 Surface Ozone, NO, and NOx Langley Data Center Data Set. It supports the NARSTO Model Intercomparison activity described in the report of a workshop that was held in Washington, DC on May 27-28, 1998. The intercomparison activity will compare meteorological, emissions, and air quality models that are used to estimate ozone concentrations in the northeastern United States. The air quality models are used to estimate how ambient ozone concentrations will change in response to changes in VOC and NOx emissions. These data are a subset of the measurements made during the NARSTO-Northeast 1995 intensive field campaign and will be used to verify model predictions. Included are surface one-hour average O3, NO, and NOx measurement results from all reporting sources for 1995. ASCII data files are available for specific time intervals and the full monitoring period. A measurement station description file is included.NARSTO, which has since disbanded, was a public/private partnership, whose membership spanned across government, utilities, industry, and academe throughout Mexico, the United States, and Canada. The primary mission was to coordinate and enhance policy-relevant scientific research and assessment of tropospheric pollution behavior; activities provide input for science-based decision-making and determination of workable, efficient, and effective strategies for local and regional air-pollution management. Data products from local, regional, and international monitoring and research programs are still available.
ATom: Data Stream for Modeling the Reactivity of ATom Air Parcels, 2016-2018
This dataset provides Modeling Data Stream (MDS) and Reactivity Data Stream (RDS) products for each of the four ATom campaigns conducted from 2016 to 2018. MDS files contain the atmospheric constituents needed to model the RDS of the air parcels along ATom flight paths. The MDS is a continuous data stream (every 10 seconds) of the atmospheric content of these key chemical species derived from the in-situ measurements collected along ATom flight paths (as reported in the comprehensive related dataset ATom: Merged Atmospheric Chemistry, Trace Gases, and Aerosols). Values for chemical species measured by multiple instruments were selected from the instrument with better coverage and/or greater precision. Missing values were filled using interpolation for short gaps. For long gaps owing to instrument failure, values were estimated using multiple linear regressions from comparable parallel flights from other ATom campaigns. All species were flagged for instrument source and values were flagged for gap-filling status. In combination, MDS and RDS provide, in essence, a photochemical climatology for each air parcel along ATom flight paths containing the reactive species that control the loss of methane and the production and loss of ozone.
CLPX-Model: Rapid Update Cycle 40km (RUC-40) Model Output Reduced Data, Version 1
The Rapid Update Cycle, version 2 at 40km (RUC-2, known to the Cold Land Processes community as RUC40) model is a Mesoscale Analysis and Prediction System (MAPS) data set that uses the Model Output Reduced Data Set (MORDS) version. This data set has been subsetted for use in the Cold Land Processes Field Experiment (CLPX).
SAFARI 2000 ETA Atmospheric Model Data, Wet and Dry Seasons 2000
With modern computer power now capable of making mesoscale model output available in real time in the operational environment, increased attention has been given to utilizing these models in order to improve the forecasting ability of meteorologists. The National Centers for Environmental Prediction (NCEP) has developed a step-mountain eta coordinate model generally known as the ETA Model.This NCEP ETA data assimilation and prediction system (see Mesinger et al., 1988; Black, 1994) has been used by the South African Weather Bureau/Service (SAWS) to provide operational regional forecast guidance since November 1993. SAWS used this model to produce the basic meteorological data for the SAFARI project. The SAWS ETA model is a hydrostatic model with a horizontal grid spacing of approximately 48 km and 38 vertical levels, with layer depths that range from 20 m in the planetary boundary layer to 2 km at 50 mb. There have been several major ETA Model upgrades at SAWS: in March 1996, August 1998, November 1999, and August 2001.
CMS: Simulated Physical-Biogeochemical Data, SABGOM Model, Gulf of America, 2005-2010
This dataset contains monthly mean ocean surface physical and biogeochemical data for the Gulf of America simulated by the South Atlantic Bight and Gulf of America (SABGOM) model on a 5-km grid from 2005 to 2010. The simulated data include ocean surface salinity, temperature, dissolved inorganic nitrogen (DIN), dissolved inorganic carbon (DIC), partial pressure of CO2 (pCO2), air-sea CO2 flux, surface currents, and primary production. The SABGOM model is a coupled physical-biogeochemical model for studying circulation and biochemical cycling for the entire Gulf of America to achieve an improved understanding of marine ecosystem variations and their relations with three-dimensional ocean circulation in a gulf-wide context.
Monthly gridded Global Land Data Assimilation System (GLDAS) from Noah-v3.3 land hydrology model for GRACE and GRACE-FO over nominal months
The total land water storage anomalies are aggregated from the Global Land Data Assimilation System (GLDAS) NOAH model. GLDAS outputs land water content by using numerous land surface models and data assimilation. For more information on the GLDAS project and model outputs please visit https://ldas.gsfc.nasa.gov/gldas. The aggregated land water anomalies (sum of soil moisture, snow, canopy water) provided here can be used for comparison against and evaluations of the observations of Gravity Recovery and Climate Experiment (GRACE) and GRACE-FO over land. The monthly anomalies are computed over the same days during each month as GRACE and GRACE-FO data, and are provided on monthly 1 degree lat/lon grids in NetCDF format. Currently, the days included in these monthly anomaly computation are same as GRACE-FO monthly Level-2 RL06.3 JPL solutions.
NACP Regional: Original Observation Data and Biosphere and Inverse Model Outputs
This data set contains the originally-submitted observation measurement data, terrestrial biosphere model output data, and inverse model simulations that various investigator teams contributed to the North American Carbon Program (NACP) Regional Synthesis activities. The data set provides nine (9) data packages of remote sensing and ground observation measurements (OM) (MODIS gross primary productivity (GPP), MODIS net primary production (NPP), MODIS fraction of photosynthetically active radiation (fPar), MODIS leaf area index (LAI), MODIS enhanced vegetation index (EVI), MODIS normalize difference vegetation index (NDVI), Forest Inventory and Analysis (FIA) forest biomass, National Agricultural Statistics Service (NASS) crop NPP, and Flux Anomaly). The data set also provides data packages of simulation results from 19 terrestrial biosphere models (TBM) and eight (8) inverse models (IM). The data packages are respectively OM, TBM, and IM data files listed in Tables 4-6. Each OM, TBM, and IM data package contains all of the original data (and documentation, if any) that the NACP Modeling and Synthesis Thematic Data Center (MAST-DC) acquired or received. These originally-submitted data were processed by the MAST-DC to produce the three standardized gridded data sets of carbon flux for inter-comparison purposes (see Related Data Products below). These original data and documentation are provided to allow users of the standardized gridded data products to be able to trace back to the data origins when needed. The Data Center (ORNL DAAC) transformed some of the originally-submitted data files to file formats that are more suitable for long-term archiving. For example, *.xlsx files were saved as *.csv, ERDAS Imagine files were converted to GeoTIFFs, and MATLAB files were converted to GeoTIFF and NetCDF formats as appropriate. Files received in NetCDF, GeoTIFF, and HDF formats were not transformed.
NACP Regional: National Greenhouse Gas Inventories and Aggregated Gridded Model Data
This data set provides two products that were derived from the recently published North American Carbon Program (NACP) Regional Synthesis 1-degree terrestrial biosphere model (TBM) and inverse model (IM) outputs (Gridded 1-deg Observation Data and Biosphere and Inverse Model Outputs, Wei et al., 2013). The first product is the aggregation of the standardized gridded 1-degree TBM and IM outputs to the Greenhouse Gas (GHG) inventory zones as defined for North America (United States, Canada, and Mexico). Depending on the data availability, the monthly/yearly Net Ecosystem Exchange (NEE), Net Primary Production (NPP), Total Vegetation Carbon (VegC), Heterotrophic Respiration (Rh), and Fire Emissions (FE) outputs from the 22 TBM and 7 IM models were aggregated from the 1-degree resolution gridded format to the inventory zones and then, further divided into Forest Lands, Crop Lands, and Other Lands sectors within each inventory zone based on the 1-km resolution GLC2000 land cover map (GLC2000, 2003).The second product is the North American national GHG inventories on the scale of inventory zones which contain estimated land-atmosphere exchange of CO2 (NEE) in forest lands, crop lands, and other lands sectors. NEE estimates were synthesized from inventory-based data on productivity, ecosystem carbon stock change, and harvested product stock change, and additional information from national-level GHG inventories of the United States, Canada, and Mexico including EPA (2011) and Environment Canada (2011).An additional summary file of annual mean NEE (2000-2006)is provided for both land sectors and reporting zones in North America and was created by combining the aggregated model output and the national GHG database and is provided. The aggregated monthly and yearly model output data and the national GHG inventories data are available in comma separated value (*.csv) format files. Also provided are detailed inventory zone spatial data as an ESRI Shapefile. Included are zone names, boundaries, and zone and land cover type area attributes. For mapping convenience, the inventory zones shapefile was merged with 1-km forest, crop, and other lands masks to create a 1-km resolution reference data file that was converted to GeoTIFF format. The GeoTIFF defines to which inventory zone and land cover type each 1-km grid cell belongs.This document provides detailed information about the content, format, and processing procedures of these two data products. Detailed descriptions of the TBMs and IMs can be found in a separate companion document: NACP Regional Synthesis - Description of Observations and Models.
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.