Livewire Data Platform: File Standards Version 1.0
This technical document is a user guide to help users of the Livewire Data Platform understand the standards and requirements for storing and sharing data on the Livewire Data Platform.
SEARCH · Engineering Papers
Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.
Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.
This technical document is a user guide to help users of the Livewire Data Platform understand the standards and requirements for storing and sharing data on the Livewire Data Platform.
This data package includes hourly meteorological and soil sensor data at eight ecohydrology monitoring sites in East River Watershed, Colorado as part of the Watershed Function Scientific Focus Area (WFSFA) research led by Lawrence Berkeley National Lab (LBNL). Four field sites were located on the hillslope of East River (ER) near Pump House (PH) at Mount Crested Butte (ER-PHS1 to 4), and the other four are in the Snodgrass Mountain (SG) area (SG-EHS5 to 8). In terms of vegetation cover, three sites are in montane grasslands (ER-PHS1, ER-PHS2, and SG-EHS5), three are below evergreen conifer canopy (ER-PHS3, SG-EHS6, and SG-EHS7), and two are below deciduous aspen canopy (ER-PHS4 and SG-EHS8). The monitoring period began in October 2019 at the East River sites, in October 2020 at SG-EHS5 and SG-EHS6, and in October 2021 at SG-EHS7 and SG-EHS8. In September 2024, all four East River sites were fully retired. The four Snodgrass Mountain sites remain active. Each site is equipped with a comprehensive suite of meteorological sensors on a tripod and soil sensors that measure weather, energy fluxes, and soil variables. This data package includes measurements from ten different types of sensors and up to thirteen individual sensors per site, including (1) a weather station (measurement height ranges from 2.8~3.8 meters (m) above ground), (2) a quantum sensor for photosynthetic active radiation (PAR) (2.4~3.3m), (3) a net radiometer (1.7~2.1m), (4) an infrared radiometer (1.6~2.2m), (5) a sonic distance sensor (1.5~1.9m), (6) a soil carbon dioxide (CO2) flux chamber (0m), (7) a soil heat flux plate (-0.05m below ground), (8) a soil oxygen sensor (-0.3m), (9) a soil water potential sensor (-0.3m), and (10) soil water content sensors at 3~4 depths (-1.15 ~ -0.1m). A total of twenty-three variables is reported in this data package, including (1) atmospheric variables: air temperature (TA), atmospheric pressure (PA), vapor pressure (VP), and vapor pressure deficit (VPD), (2) precipitation variables: rain precipitation (P) and snow depth (D_SNOW), (3) energy fluxes variables: four-component net radiation (NETRAD) (shortwave/longwave incoming/outgoing radiation, SW_IN, SW_OUT, LW_IN, LW_OUT), photosynthetic photon flux density (PPFD), and soil heat flux (G), (4) soil variables: soil water content (SWC), soil water potential (SWP), soil temperature (TS), soil bulk electrical conductivity (COND_SOIL), and soil gaseous oxygen concentration (O2_SOIL), (5) wind variables: two-dimensional wind speed (WS), gust speed (WS_MAX), and wind direction (WD), and (6) surface variables: surface infrared temperature (T_CANOPY) and soil CO2 flux (CO2_SOIL). Please see the Methods section for data processing and QA/QC steps taken to generate the hourly datasets. The following files are included in this data package (notes on version: v{x}-{y}, where x is the metadata version, and y is the data version, when applicable): (1) “metadata_site_v{x}-{y}.csv” - a site metadata file that summarizes location information of all sites, including site ID, description, coordinates, timeframe, elevation, and vegetation cover, (2) “metadata_instrument_v{x}-{y}.csv” - an instrument metadata file that summarizes sensor information of all sites, including sensor manufacturer and model, measurement height, and sampling and averaging interval of all variables, (3) "data_{SITE_ID}_v{x}-{y}.csv" - eight data files that contain hourly data of each site indicated by {SITE_ID} in the filename, (4) “/figure/data_{SITE_ID}_v{x}-{y}.png" - eight figures that help visualize data of each site indicated by {SITE_ID} in the filename, (5) “/photo/*” - photos of each site indicated by {SITE_ID} in the filename, and (6) four file level metadata (flmd.csv) and data dictionary (*_dd.csv) files that summarize file, header, column, and variable information of all files. Notes: (1) Measurement height: Each variable name is followed by conventional positional qualifiers “H_V_R”, where H indicates the relative horizontal positions of that specific variable, V the vertical positions, and R the replicates. In this data package, only the vertical qualifier V varies, and V increases from the highest vertical position (V=1) to the lowest. Variables with the same qualifier are not necessarily measured by the same sensor, and the same variable with the same qualifier across different sites are not necessarily measured at the same height. Please refer to “metadata_instrument.csv” for the sensor information and measurement heights, and whether a variable is measured below the canopy. (2) Variable availability: Snow depth is not available at ER-PHS3 and SG-EHS7. SWC, soil temperature, and soil bulk EC at the deepest depth (<-1m) are not available at SG-EHS6 and SG-EHS7. The missing value code for numeric variables is -9999, except for SWP. For SWP, the missing value code is +9999, because SWP values are negative. (3) Sampling frequency: Please refer to “metadata_instrument.csv” for the increase of sampling frequency of some variables from 30-min to 1-min at ER-PHS1 to 4 in July 2020. (4) Sensors: While the methods of each sensor are not detailed, all sensors are commercially available, and their methods can be found in their manuals. Please refer to “metadata_instrument.csv” for the sensor manufacturer and model information. This work was supported by the Watershed Function Science Focus Area at Lawrence Berkeley National Laboratory funded by the US Department of Energy, Office of Science, Biological and Environmental Research under Contract No. DE-AC02-05CH11231.
Abstract. A new weakly coupled land data assimilation (WCLDA) system based on the four-dimensional ensemble variational (4DEnVar) method is developed and applied to the fully coupled Energy Exascale Earth System Model version 2 (E3SMv2). The dimension-reduced projection four-dimensional variational (DRP-4DVar) method is employed to implement 4DVar using the ensemble technique instead of the adjoint technique. With an interest in providing initial conditions for decadal climate predictions, monthly mean anomalies of soil moisture and temperature from the Global Land Data Assimilation System (GLDAS) reanalysis from 1980 to 2016 are assimilated into the land component of E3SMv2 within the coupled modeling framework with a 1-month assimilation window. The coupled assimilation experiment is evaluated using multiple metrics, including the cost function, assimilation efficiency index, correlation, root-mean-square error (RMSE), and bias, and compared with a control simulation without land data assimilation. The WCLDA system yields improved simulation of soil moisture and temperature compared with the control simulation, with improvements found throughout the soil layers and in many regions of the global land. In terms of both soil moisture and temperature, the assimilation experiment outperforms the control simulation with reduced RMSE and higher temporal correlation in many regions, especially in South America, central Africa, Australia, and large parts of Eurasia. Furthermore, significant improvements are also found in reproducing the time evolution of the 2012 US Midwest drought, highlighting the crucial role of land surface in drought lifecycle. The WCLDA system is intended to be a foundational resource for research to investigate land-derived climate predictability.
We have performed an initial assessment of available data libraries for photonuclear reactions on 239 Pu. Specifically, we considered the following five evaluated data libraries: the ENDF/B-VIII.0 and the recently deployed ENDF/B.VIII.1 Evaluated Nuclear Data Files, the International Atomic Energy Agency (IAEA) Photonuclear Data Library 2019 (IAEA-2019), the Japanese evaluated nuclear data library version 5 (JENDL-5), and the 2023 version of the TALYS Evaluated Nuclear Data Library (TENDL-2023). Each of these libraries were translated form ENDF to GNDS format using FUDGE. For each of the Pu isotopes, we provide a comparison with available experimental data in the EXFOR experimental nuclear reaction database, as well as a rapid quantitative assessment of the quality of the evaluation.
As of May 12, 2026 this dataset is currently being versioned to include data up to April 2026. Once the versioning process is complete, new data files will be available for access. The dataset metadata will also be updated to reflect the data availability of the new data being versioned. This dataset was collected at the UIC Plant Research Laboratory in Chicago, Illinois, as part of the Community Research on Climate and Urban Science (CROCUS) Urban Integrated Field Laboratory (UIFL) project, led by Argonne National Laboratory. The site provides continuous atmospheric flux measurements, focusing on CO₂, H₂O, and heat and momentum transport in an urban setting. The data is processed at 30 minutes interval using the Eddy Covariance method and includes quality control and diagnostic data generated by EddyPro software. The data is stored in the netCDF files following CF conventions. The UIC Plant Research Laboratory is located near major highways and urban infrastructure, including buildings and parking areas. The surrounding landscape consists of a mix of turf, plants, trees, and impervious surfaces such as concrete and asphalt, making it ideal for studying urban at for studies on urban sustainability, air quality, and the effects of urbanization on atmospheric processes on urban climate dynamics, air quality, and surface-atmosphere exchanges within the city of Chicago. This dataset is funded by the U.S. Department of Energy’s Office of Science, Biological and Environmental Research (BER) program.
132 urban parameters for Maricopa County are provided in WRF-readable binary form at 100m grid spacing with index. These parameters allow regional weather modelers to account for the contribution of built structures in modeling the urban heat island. This data is also provided in CSV format. Data are generated using 2023 Model America Version 1 data (https://www.osti.gov/biblio/1774134) and will be used to represent the urban morphology in Maricopa County to the WRF model for experiments conducted for the Southwest Urban Corridor Integrated Field Laboratory (SW-IFL). Three files associated with these parameters are included in this repository: index, binary file 00001-01549.00001-01796, Maricopa_Parameters_03072024.csv . To use the binary file with the Weather Research and Forecasting (WRF) model, the binary file and the index file must be placed in their own directory in WRF_GEOG and accessed in the same way the National Urban Database and Access Portal Tool data (NUDAPT44) would be accessed. The name of the binary file must be in the format: 00001-02715.00001-01803. If, upon download, the binary filename has additional or different characters, please change them to match this format.
As of April 9, 2026: this dataset is currently being versioned to include data up to March 31, 2026. Once the versioning process is complete, new data files will be available for access. The dataset metadata will also be updated to reflect the data availability of the new data being versioned. Raw atmospheric measurements collected at the University of Illinois Chicago (UIC) as part of the Community Research on Climate and Urban Science (CROCUS) project. The data includes high-frequency measurements of carbon dioxide (CO₂) concentration, water vapor (H₂O) concentration, wind speed (U, V, W components), temperature, and atmospheric pressure. These raw data are recorded by the LI-7500DS Open Path CO₂/H₂O Analyzer and a sonic anemometer at a 10 Hz acquisition frequency, providing the necessary inputs for calculating fluxes of CO₂, H₂O, heat, and momentum. In addition to gas concentration measurements, the dataset includes diagnostic information from the instruments, including absorptance, sample and reference signals, and diagnostic values. The data were collected to study urban atmospheric conditions and contribute to flux calculations for urban climate research. These measurements form the basis for calculating 30-minute average fluxes of key atmospheric variables using the eddy covariance method. This dataset provides critical raw input data for researchers interested in atmospheric fluxes, urban air quality, and the interaction between the urban environment and atmospheric processes. The data is part of the U.S. Department of Energy’s Biological and Environmental Research (BER) program, under the CROCUS Urban Integrated Field Laboratory (UIFL) project.
This data package contains root trait data collected from 170 plots of rapidly expanding shrub genera (Alnus, Betula, and Salix) and a widespread sedge (Eriophorum vaginatum) along a latitudinal and temperature gradient in northern Alaska. The trait data were collected in July 2017 and include root architecture (root diameter and branching patterns), mycorrhizal colonization (%), nitrogen concentration (%), delta 15N (per mil), and vertical root biomass. These raw data support a submitted manuscript that examines the distribution and interspecific variations of absorptive root traits of shrubs and graminoids across the graminoid-dominated nutrient-poor arctic tundra and reveals how deciduous shrub expansion affects plant nutrient acquisition strategies in tundra ecosystems. Data are presented by site (n=5) and patch (shrub or sedge plot) in separate csv files. The location data are provided in the “plot coordinates” file; all other files contain the data in the file title. Detailed methods are in Chen et al. (2020).2023/07/20 Update: The latest version of this dataset publication is version 2.0. The latest version of the data package was updated to include the alder nodule biomass dataset (alder nodule biomass.csv) and associated metadata (metadata_alder nodule biomass.csv). The name of the previous metadata file was updated (metadata_ root traits and biomass.csv ) to distinguish it from the new metadata file.
This dataset contains gridded monthly Leaf Area Index (LAI) derived from the daily NOAA Climate Data Record (CDR) of AVHRR (Version 5) and VIIRS (Version 1) Leaf Area Index (LAI) and Fraction of Absorbed Photosynthetically Active Radiation (FAPAR). This data record spans from 1981 to 2024 using data from NOAA polar orbiting satellites: NOAA-7, -9, -11, -14, -16, -17, -18, -19 and S-NPP. The data are projected on a 0.05 degree x 0.05 degree global grid, as in the original CDR. The original CDR is one of the Land Surface CDR products produced by the NASA Goddard Space Flight Center (GSFC) and the University of Maryland (UMD), which is accompanied by algorithm documentation, data flow diagram and source code for the NOAA CDR Program. This dataset is in the netCDF-4 file format following ACDD and CF Conventions. This dataset has applied quality assurance information to only include "OK" data from the original CDR in the monthly aggregation.
This data package contains leaf and size trait and environmental data collected from 170 plots of rapidly expanding shrub genera (Alnus, Betula, and Salix) and a widespread sedge (Eriophorum vaginatum) along a latitudinal and climate gradient in northern Alaska. The trait data were collected in summer 2017 and include leaf area, specific leaf area, leaf nitrogen concentration, leaf delta 15N, leaf delta 13C, shrub height, aboveground biomass, and the ratio of root biomass to aboveground biomass. These raw data support a submitted manuscript that examines the intraspecific variations of size, leaf and root traits of shrubs across the graminoid-dominated nutrient-poor arctic tundra and reveals the environmental drivers of deciduous shrub traits in tundra ecosystems (Fraterrigo et al., submitted). Data are presented by site (n=5) and patch (shrub or sedge plot) in separate csv files. The location data are provided in the “plot coordinates” file; all other files contain the data in the file title. Detailed methods are in the submitted manuscript. A companion data package contains root trait data collected simultaneously from the same plots (Fraterrigo and Chen, 2020). See the "Related references" section for more information.2023/07/20 Update: The latest version of this dataset publication is version 2.0. The latest version of the data package was updated to correct the leaf size trait.csv file.
Statement of purpose: Cyclones alter the function and composition of tropical forests, making effects of intensifying cyclones on carbon-rich forests a critical topic of study. Here, we quantified cyclone-induced damage and recovery of 21 cyclone disturbances affecting 23 pantropical forest sites between 1988-2017 utilizing leaf area index (LAI), enhanced vegetation index (EVI), normalized difference vegetation index (NDVI), and transformed NDVI (kNDVI) values from Google Earth Engine. Field observations collected in a meta-analysis (Bomfim et al., 2022, in review) were used to ground-truth and test effects of soil resource availability and disturbance factors on damage and recovery. This meta-analysis also served as the basis to begin vegetation index extraction, utilizing unique site and date combinations, from tropical forests effect by cyclone disturbances. We began collecting NDVI (5km resolution) from the NOAA Climate Data Record (CDR) of AVHRR Normalized Difference Vegetation Index (NDVI), Version 5 data product (Vermote, 2019) for all case studies included, 42. Next, we began extracting Landsat data from Landsat 4, 5, and 8, courtesy of the U.S. Geological Survey, in search of higher resolution data. We selected a 3 by 3 Landsat pixel area, leading to a 90m resolution data extraction. The specific imagery used includes Landsat 4 USGS Landsat 4 TM Collection 1 Tier 1 TOA (top of atmosphere) Reflectance, Landsat 5 USGS Landsat 5 TM (thematic mapper) Collection 1 Tier 1 TOA Reflectance, and Landsat 8 USGS Landsat 8 Collection 1 Tier 1 TOA Reflectance. Within Google Earth Engine, we selected the date and location (latitude and longitude), calculated NDVI, kNDVI, and EVI utilizing Landsat bands (see metadata_NGEE-tropics_cyclones), and extracted post- and pre-cyclone values for each case study to calculate cyclone-induced change in the vegetative index. Due to limited spatial resolution of Landsat remote sensing data, MODIS products were investigated next. First, the MOD13Q1.006 Terra Vegetation Indices 16-Day Global 250m product was used to extract 250m EVI and NDVI (Didan, 2015) and then the MCD15A3H.006 MODIS Leaf Area Index/FPAR 4-Day Global 500m product product was used to extract LAI 500m (Myneni et al., 2015). Pre- and post-cyclone values, change in the vegetative index, and standard deviation for all values are included in the main csv (see case_study_data.csv) for all vegetative indices collected, including LAI 500m, EVI 250m, NDVI 250m, NDVI 90m, kNDVI 90m, EVI 90m, and NDVI 5km. Lastly, recovery values were calculated utilizing a standardization method (see metadata_NGEE-tropics_cyclones) and recovery values for MODIS (see MODIS_recovery.csv) and Landsat (Landsat_recovery.csv) data are included.
This data package is associated with the publication “Accuracy evaluation of cost-effective 3D reconstruction approaches for hydrobiogeochemical processes in non-perennial stream riverbeds” published in Frontiers in Environmental Science, Environmental Informatics and Remote Sensing (Bao et al., 2026; doi: 10.3389/fenvs.2026.1725258). This data package includes the drone photos for a section of Umtanum Creek in Washington, Unted States. The photos were used to reconstruct the 3-dimensional (3D) digital elevation model (DEM) of the riverbed for the investigated stream section. The reconstruction results from four approaches are provided: (1) unoccupied aerial vehicle (UAV, colloquially known as drone) imagery-based Structure-from-Motion (SfM), (2) a machine learning-based 3D reconstruction model, Visual Geometry Grounded Deep Structure from Motion (VGGSfM), (3) Visual Geometry Grounded Transformer for long sequence of images (VGGT-Long), and (4) handheld smartphone LiDAR scanning. The ground truth measurements by tripod-mounted optical level kit and ground control points GPS locations for evaluating the accuracy of the four reconstruction approaches are also provided in this data package. A preliminary version of this data package was published in October 2025 at the time of manuscript submission. It was updated in March 2026, at the time of manuscript acceptance, to include additional metadata (this readme, data dictionary, and file level metadata). The data did not change. For details on how to navigate data packages generated by this project, see https://data.ess-dive.lbl.gov/portals/PNNLRiverCorridorSFA/About. In addition to a readme, this data package also includes a file-level metadata (FLMD) file that describes each file and a data dictionary (DD) that describes all column/row headers and variable definitions. This dataset is comprised of (1) 8 folders; (2) the detailed flight configuration html files; (3) field metadata; (4) a readme; (5) a data dictionary; and (6) file-level metadata. The folders “2024_10_18_d01” and “2024_10_18_d02” contain the original drone photos for the two drone flights (d01 and d02) on October 18, 2024. The reconstruction results from each of the approaches are in the folders called “ODM_SfM”, “VGGSfM”, “VGGTLong”, and “LiDAR”. The ground truth measurements are in the folder called “optical_level_kit”. Lastly, results comparing the different approaches are in the folder called “comparisons”. All files are .csv, .html, .jpg, .obj, .txt, and .npy. For information on using the .obj and .npy files, see the readme files within the same folder as the files.
This dataset contains derived data from the profiling lidar (Windcube v2.1) deployed at WFIP3's BARG. The derived data files herein are based on the lidar's STA files (i.e., the 10-min average data files). Note: this version of the data has been corrected to account for the motion of the barge.
With a potentially increasing share of the electricity grid relying on wind to provide generating capacity and energy, there is an expanding global need for historically accurate, spatiotemporally continuous, high-resolution wind data. Conventional downscaling methods for generating these data based on numerical weather prediction have a high computational burden and require extensive tuning for historical accuracy. In this work, we present a novel deep learning-based spatiotemporal downscaling method using generative adversarial networks (GANs) for generating historically accurate high-resolution wind resource data from the European Centre for Medium-Range Weather Forecasting Reanalysis version 5 data (ERA5). In contrast to previous approaches, which used coarsened high-resolution data as low-resolution training data, we use true low-resolution simulation outputs. We show that by training a GAN model with ERA5 as the low-resolution input and Wind Integration National Dataset Toolkit (WTK) data as the high-resolution target, we achieved results comparable in historical accuracy and spatiotemporal variability to conventional dynamical downscaling. This GAN-based downscaling method additionally reduces computational costs over dynamical downscaling by two orders of magnitude. We applied this approach to downscale 30 km, hourly ERA5 data to 2 km, 5 min wind data for January 2000 through December 2023 at multiple hub heights over Ukraine, Moldova, and part of Romania. With WTK coverage limited to North America from 2007–2013, this is a significant spatiotemporal generalization. The geographic extent centered on Ukraine was motivated by stakeholders and energy-planning needs to rebuild the Ukrainian power grid in a decentralized manner. This 24-year data record is the first member of the super-resolution for renewable energy resource data with wind from the reanalysis data dataset (Sup3rWind).
This repository contains data for version 1 of the Genome Resolved Open Watersheds database (GROWdb), along with data for the manuscript describing GROWdb.
From 24 October to 11 December 2024, the NSF-DOE Vera C. Rubin Observatory conducted an on-sky campaign using the engineering LSST Commissioning Camera (LSSTComCam) to test the end-to-end functionality of hardware and software, as well as operational procedures. This interim report provides a preliminary technical overview of our understanding of the integrated system performance based tests and analyses conducted during the LSSTComCam on-sky campaign. The objectives are to synthesize what we have learned about the system in a timely way to inform on-going commissioning efforts, and to inform the Rubin science community on the progress of the LSSTComCam on-sky campaign. The report is organized into sections to describe major activities during the campaign, as well as multiple aspects of the demonstrated system and science performance. All of the results presented here are to be understood as work in progress using engineering data and the initial versions of the data processing pipelines; the report is a living document that will be updated as analyses are refined.
FIDASIM is a code that models signals produced by charge-exchange reactions between neutrals and ions (both fast and thermal) in magnetically confined plasmas. With the ion distribution function as input, the code predicts the efflux to a neutral particle analyzer diagnostic and the photon radiance of Balmer-alpha light to a fast-ion D α diagnostic, in addition to many other related quantities. A new, parallelized version of the Monte Carlo code FIDASIM has been developed in Fortran90 that is substantially faster than the original interactive data language version. Modified algorithms include more accurate treatments of the time dependent collisional-radiative equations that describe neutral energy levels, of the cloud of ‘halo’ neutrals that surround the injected neutral beam, and of finite Larmor radius effects. Enhanced physics capabilities include modelling ‘passive’ signals from cold edge neutrals, the ability to treat general three-dimensional magnetic confinement configurations, and calculations of diagnostic-specific weight functions that enable tomographic reconstructions of the fast-ion distribution function. Neutral beam attenuation, beam emission, and fast-ion birth profiles are also modelled. Finally, the new algorithms have been successfully validated against experimental data and new features have been tested through benchmarks between two independently developed versions of the code.
The most recent version of these data are available https://doi.org/10.25581/spruce.100/1874948 and supersedes all previous versions. This data set consists of PhenoCam data from the SPRUCE experiment from the beginning of whole ecosystem warming (Hanson et al. 2017) in August 2015 through March 31 of 2021, with start- and end-of-season phenological transition dates derived through the end of autumn 2020. Digital cameras, or phenocams, installed in each SPRUCE enclosure track seasonal variation in vegetation “greenness”, a proxy for vegetation phenology and associated physiological activity. Three separate regions of interest (ROIs) were defined for each camera field of view, corresponding to different vegetation types and demarcating (1) Picea trees (vegetation type EN, for evergreen needleleaf); (2) Larix trees (vegetation type DN, for deciduous needleleaf); and (3) the mixed shrub layer (vegetation type SH). User note: A list of previous versions can be found in the Related Datasets section of the user guide.