Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Data reporting format”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Association Kinetics for Perfluorinated n -Alkyl Radicals

Radical-radical reaction channels are important in the pyrolysis and oxidation chemistry of perfluoroalkyl substances (PFAS). In particular, unimolecular dissociation reactions within unbranched n-perfluoroalkyl chains, and their corresponding reverse barrierless association reactions, are expected to be significant contributors to the gas-phase thermal decomposition of families of species such as perfluorinated carboxylic acids and perfluorinated sulfonic acids. Unfortunately, experimental data for these reactions are scarce and uncertain. Furthermore, obtaining reliable theoretical predictions for such reactions is a laborious and computationally intensive task. Here, in this work, the chemical kinetics of the various association/decomposition reactions producing/decomposing the C 2 -C 4 series of unbranched n-perfluoroalkanes (C 2 F 6 , C 3 F 8 , and C 4 F 10 ) are examined using state-of-the-art ab initio transition-state-theory-based master-equation calculations. The variable-reaction-coordinate transition-state theory (VRC-TST) formalism is employed in computing the microcanonical and canonical rates for the association reactions. Reaction thermochemistry is obtained via composite quantum chemistry calculations and the laddering of error-canceling reaction schemes via a connectivity-based hierarchy approach employing ANL1/ANL0-style reference energies. Lennard-Jones collision model parameters for the considered systems were estimated by a direct dynamics approach, and collisional energy transfer parameters were obtained from analogies to systems of similar size and heavy-atom connectivity. A one-dimensional master equation approach was used to convert the microcanonical rate coefficients from the VRC-TST analysis into temperature- and pressure-dependent rate constants for the association reactions and the reverse dissociation reactions. The data are reported in standardized formats for usage in comprehensive chemical kinetic models for PFAS thermal destruction.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Estimating soil thermal inertia profiles from the passive equilibration of a temperature probe: Supporting data and code

This dataset contains measurements of soil thermal properties and organic layer thickness at three field sites on the Seward Peninsula near Nome, Alaska, USA. Code is also provided for estimation of soil thermal properties from temperature data, including a forward heat transfer mode, a library of simulated temperature data, and an inversion scheme to retrieve thermal properties from temperature measurements (Lamb et al., 2025). This dataset and code is published in support of the manuscript “Estimating Soil Thermal Inertia Profiles from the Passive Equilibration of a Temperature Probe” by Lamb et al., 2025, in which an experimental thermal measurements technique is developed. The dataset contains measurements of soil thermal properties and organic layer thickness obtained with an industry standard thermal properties analyzer and direct observations, as well as data generated using the experimental technique to estimate thermal properties and soil structure from temperature observations. All data is reported in .csv format, and all code is written in Python 3.7 using standard packages. This research was performed as a part of the NGEE Arctic project, which aims to advance model predictions of arctic carbon cycle responses to a changing climate over the 21st Century.The Next-Generation Ecosystem Experiments: Arctic (NGEE Arctic), was a research effort to reduce uncertainty in Earth System Models by developing a predictive understanding of carbon-rich Arctic ecosystems and feedbacks to climate. NGEE Arctic was supported by the Department of Energy's Office of Biological and Environmental Research. The NGEE Arctic project had two field research sites: 1) located within the Arctic polygonal tundra coastal region on the Barrow Environmental Observatory (BEO) and the North Slope near Utqiagvik (Barrow), Alaska and 2) multiple areas on the discontinuous permafrost region of the Seward Peninsula north of Nome, Alaska. Through observations, experiments, and synthesis with existing datasets, NGEE Arctic provided an enhanced knowledge base for multi-scale modeling and contributed to improved process representation at global pan-Arctic scales within the Department of Energy's Earth system Model (the Energy Exascale Earth System Model, or E3SM), and specifically within the E3SM Land Model component (ELM).

54 ENVIRONMENTAL SCIENCES↗

Tussock tundra surface temperatures, ambient air and incoming photosynthetically active radiation measured at the NGEE Arctic Council site, 2021 - 2023

This dataset contains temperature measurements carried out along two fiber optics cables/lines (150 m each) laid out along the ground at the Next Generation Ecosystem Experiment (NGEE) Arctic site near Council, Alaska. The lines traverse an heterogeneous part of the tussock tundra site including thermokarst features and lichen dominated sections of the tundra. Measurements were done using a Sensornet Oryx DTS, a Distributed Temperature sensor that was installed in September 2021 and taken down in August 2023. The sensor was powered by solar power with data being collected every 30 minutes at 1 m resolution. In addition to these measurements air temperature and incoming photosynthetically active radiation (PAR) are provided. These measurements are co-located with the NGEE Arctic Council eddy flux and meteorological station (AmeriFlux ID US-NGC). Included are six *.csv files (four data files and two reporting format files) and two *.kml files. The Next-Generation Ecosystem Experiments: Arctic (NGEE Arctic), was a research effort to reduce uncertainty in Earth System Models by developing a predictive understanding of carbon-rich Arctic ecosystems and feedbacks to climate. NGEE Arctic was supported by the Department of Energy's Office of Biological and Environmental Research. The NGEE Arctic project had two field research sites: 1) located within the Arctic polygonal tundra coastal region on the Barrow Environmental Observatory (BEO) and the North Slope near Utqiagvik (Barrow), Alaska and 2) multiple areas on the discontinuous permafrost region of the Seward Peninsula north of Nome, Alaska. Through observations, experiments, and synthesis with existing datasets, NGEE Arctic provided an enhanced knowledge base for multi-scale modeling and contributed to improved process representation at global pan-Arctic scales within the Department of Energy's Earth system Model (the Energy Exascale Earth System Model, or E3SM), and specifically within the E3SM Land Model component (ELM).

Air temperature↗

LTE-P-19 Comparison Metrics and Terms for Low Temperature Electrolysis

Defines a standard set of definitions, metrics, units, and conventions for low temperature electrolysis (LTE). The goal is to ensure that the literature is consistent from research group to research group, and data is reported in similar formats, such that results can be compared on a similar basis.

Electryolyzer↗

A Non-Invasive Approach for Elucidating the Spatial Distribution of In Situ Stress in Deep Subsurface Geologic Formations Considered for CO 2 Storage (Task 2 Report Extracting Stress Data from Seismic Data)

This report describes research accomplishments achieved with funding provided through Department of Energy (DOE) Contract DE-FE0031686. The purpose of this DOE funding is to develop methods that can provide key information about stress fields that act on deep rocks without invading the earth to acquire that information. The research objective was to demonstrate methods that extract the azimuth directions of SHmax and SHmin stress in deep rocks from traditional seismic reflection data like the data that are used to explore for deep oil and gas reservoirs. Battelle performed two seismic investigations and achieved estimates of SHmax and SHmin azimuths in both efforts that agreed with local, non-seismic, ground-truth measurements of SHmax orientation. The first procedure utilized vertical seismic profiling (VSP) data; the second effort utilized three-dimensional (3D) seismic data.

58 GEOSCIENCES↗

ESS-DIVE Reporting Format for Location Metadata

The ESS-DIVE location metadata reporting format provides instructions and templates for reporting a minimum set of metadata for discrete point locations in geographic space represented by x, y, and z coordinates. This format was created based on a need for earth and environmental science researchers to more consistently provide metadata about locations where they conduct studies. To create the format, we incorporated elements from ESS-DIVE’s community reporting formats as well as 12 additional data standards or other data resources (e.g., databases, data systems, or repositories). In the template, we ask researchers to indicate unique locations using Location IDs and indicate hierarchies of locations through parent location IDs. We also provide additional optional fields for researchers to indicate how they measured the point location and the date and time that the location was first used as a research siteThis dataset contains support documentation for the reporting format (README.md and instructions.md), a terminology guide (guide.md), a crosswalk indicating how this reporting format relates to existing standards and data resources (Location_metadata_crosswalk.csv), a data dictionary (dd.csv), file-level metadata (flmd.csv), and the location metadata templates in both CSV (Location_metadata_template.csv) and Excel formats (Location_metadata_template.xlsx).

54 ENVIRONMENTAL SCIENCES↗

Leaf gas exchange, leaf water potential and spectral reflectance, BIONTE, Brazil, 2023

Leaf traits measured at the top of the canopy on 30 tree species in the BIONTE experimental forest, near Manaus, Brazil, at three time points during August to November, 2023. The aim of this study was to test the relationship between leaf water use efficiency and wood density. Leaf gas exchange was measured on top-of-canopy leaves accessed by an articulated boom lift. Response curves (A-Ci and A-Q), and dark adapted respiration were measured on cut branches. Survey gas exchange measurements were performed on attached leaves at various times during the day. Following gas exchange, leaves were harvested, and measured for leaf water potential and spectral reflectance. Data in this data package is presented in csv files, with metadata in csv and descriptive experimental protocols as pdf. Gas exchange data and metadata meets the requirements of the ESS-DIVE reporting format for leaf-level gas exchange data and metadata.

54 ENVIRONMENTAL SCIENCES↗

Soil microbial ecology and microbiome-metabolite linkages improve understanding of ecosystem states along terrestrial-aquatic interfaces

These data are from Bandopadhyay et al., "Soil microbial ecology and microbiome-metabolite linkages improve understanding of ecosystem states along terrestrial-aquatic interfaces". This study aims to understand the soil microbial ecology along terrestrial-aquatic interfaces of a freshwater and estuarine region and how it relates to organic matter. We analyzed soil microbial (16S rRNA gene) and organic matter (Fourier-transform ion cyclotron resonance mass spectrometry, FTICR-MS) composition from upland (forested), transition (stressed forest), and wetland positions at three sites in each of the Lake Erie (freshwater) and Chesapeake Bay (estuarine) regions. This dataset includes 16S rRNA gene amplicon data (only processed file types included here) and organic matter composition from FTICR-MS data (raw and processed files included here) from upland (forested), transition (stressed forest), and wetland positions at three sites in each of the Lake Erie and Chesapeake Bay regions. These sites are part of the COMPASS-FME project (https://compass.pnnl.gov/FME/COMPASSFME). File formats and software needed to access files: 16S rRNA gene amplicon data: These files follow the format reported here https://ess-dive.gitbook.io/amplicon-sequencing-reporting-format#updates-in-v1.0.1. As per this format, there are four file types reported: 1. Taxon tables (also called sequence-by-sample or OTU (operational taxonomic unit)/ESV (exact sequence variant) tables) : available in a .txt file format and accessible using TextEdit or MS Excel. 2. Representative sequences (also called consensus sequences) : available in a .fasta format and accessible using TextEdit. 3. Sequencing metadata : available in a MS Excel workbook file format and CSV file format 4. Bioinformatic metadata : available in a MS Excel workbook file format and CSV file format FTICR-MS data: 1. Raw data converted to a processed file with intensities of the peaks in the given samples : available in a MS Excel CSV file format 2. Processed file used in analyses and visualizations (appended as icr_long_) : available in a MS Excel CSV file format 3. Metadata file for ICR features (appended as icr_meta) : available in a MS Excel CSV file format

54 ENVIRONMENTAL SCIENCES↗

Leaf gas exchange and fitted parameters, two sites in Panama, 2022

Photosynthetic CO2 response curves (ACi curves), light response curves (AQ curves), dark adapted dark respiration, conductance curves, and survey measurements for leaves measured in the Parque Natural Metropolitano (PNM) and Gamboa, Panama, from January to April 2022 are presented. Measurements were made on leaves from 46 different tree, shrub and liana species, from sunlit canopy and understory locations. Leaf area index from 8 vertical profiles at PNM are also included. The aim of this measurement campaign was three-fold: (i) to improve our understanding of the vertical variation in leaf-level water use efficiency; (ii) to develop an understanding of the sensitivity of stomata to changes in environmental conditions, especially light, humidity, and temperature; (iii) to improve models which can predict leaf traits from leaf contact spectral measurements. All the gas exchange data and metadata are presented in .csv files and complete instrument output are included in .zip folders. Data and metadata meet the ESS-DIVE leaf-level gas exchange reporting format requirements. The protocol details are provided as pdf documents. In addition to gas exchange data reported here these samples were also used for measurement of leaf optical properties, carbon and nitrogen content, and leaf mass per unit leaf area (LMA). These data can be cross linked using the unique sample ID and are provided in separate data packages.

54 ENVIRONMENTAL SCIENCES↗

Groundwater and river water elevations and temperature from 2017 to 2022 across Meander Z in the East River Watershed, Colorado

This dataset includes groundwater and river water elevations and temperature data collected in the East River watershed located in the Upper Colorado River Basin. The data were collected in order to investigate the coupling between hydrology and biogeochemical processes in the floodplain. Data was collected at ten groundwater locations in Meander Z (MZ), located just upstream of the confluence with Brush Creek and two river locations directly adjacent to Meander Z from 2017-2019. From 2019-2022, data was collected at five groundwater locations in Meander Z. Note that location names, not location identifiers (IDs), are used in the related publication Dewey et al. (2022). Both location IDs and names are included in data files. Files in this dataset include the main data files for each location zipped into a single folder (waterlevel_data.zip), an installation methods file describing sensor installation (InstallationMethods.csv), a file containing field metadata including GPS (Global Positioning System) coordinates and ground surface elevations (transducers_locations.csv). This dataset also includes a file-level metadata (flmd.csv) file that lists each file contained in the dataset with associated metadata and a data dictionary (dd.csv) file that contains column/row headers used throughout the files along with a definition, units, and data type. This dataset conforms to the ESS-DIVE hydrological reporting format. 2026-04-27 Update: The river water elevation data files (ER-MZR1.csv and ER-MZR2.csv) were corrected. The data for these two locations were inadvertently swapped in the original published data. This work was supported by the Watershed Function Science Focus Area at Lawrence Berkeley National Laboratory funded by the US Department of Energy, Office of Science, Biological and Environmental Research under Contract No. DE-AC02-05CH11231.

54 ENVIRONMENTAL SCIENCES↗

ESS-DIVE Reporting Format for File-level Metadata

The ESS-DIVE reporting format for file-level metadata (FLMD) provides granular information at the data file level to describe the contents, scope, and structure of the data file to enable comparison of data files within a data package. The FLMD are fully consistent with and augment the metadata collected at the data package level. We developed the FLMD template based on a review of a small number of existing FLMD in use at other agencies and repositories with valuable input from the Environmental Systems Science (ESS) Community. Also included is a template for a CSV Data Dictionary where users can provide file-level information about the contents of a CSV data file (e.g., define column names, provide units). Files are in .csv, .xlsx, and .md. Templates are in both .csv and .xlsx (open with e.g. Microsoft Excel, LibreOffice, or Google Sheets). Open the .md files by downloading and using a text editor (e.g. Notepad or TextEdit). Though we provide Excel templates for the file-level metadata reporting format, our instructions encourage users to 'Save the FLMD template as a CSV following the CSV Reporting Format guidance'. In addition, we developed the ESS-DIVE File Level Metadata Extractor which is a lightweight python script that can extract some FLMD fields following the recommended FLMD format and structure.

54 ENVIRONMENTAL SCIENCES↗

Computed Tomography Scanning and Geophysical Measurements of Appalachian Basin Core from the Jones and Laughlin #1 Well, Beaver County, PA

The computed tomography (CT) facilities and the Multi-Sensor Core Logger (MSCL) at the National Energy Technology Laboratory (NETL) in Morgantown, West Virginia, were used to characterize Appalachian Basin core from Beaver County, Pennsylvania. The primary impetus of this work is a collaboration between the U.S. Department of Energy (DOE) and the Pennsylvania Geological Survey to characterize and make publicly available core information from the Onondaga-Huntersville formations of the Appalachian Basin. This stratigraphic well and the core data produced in this report will aid in understanding the structural complexities of the Onondaga-Huntersville formations. The resultant datasets are presented in this report and can be accessed from NETL's Energy Data eXchange (EDX) online system using the following link: https://edx.netl.doe.gov/dataset/jonesandlaughlin1well. All equipment and techniques used were non-destructive, enabling future examinations and analyses to be performed on these cores. Fractures, discontinuities, and millimeter-scale features were readily detectable with imaging performed with the NETL medical CT scanner over the entire core. Qualitative analysis of the medical CT images, coupled with X-ray fluorescence (XRF), and magnetic susceptibility measurements from the MSCL were useful in identifying zones of interest for further study. Targeted higher resolution CT scanning of select sections was performed with NETL’s micro-CT scanner. The combination of methods used provides a multiscale analysis of the core; the resulting macro and micro descriptions are relevant to many subsurface energy related examinations traditionally performed at NETL.

58 GEOSCIENCES↗

Leveraging large language models to address data scarcity in machine learning for graphene synthesis

Machine learning in experimental materials science faces significant challenges due to the scarcity of data, which are costly and time-consuming to generate, particularly when relying on in-house experiments. Literature data mining offers a potential solution but introduces issues like mixed data quality, inconsistent formats, and non-uniform reporting of synthesis parameters, resulting in partially missing and heterogeneous features across the dataset. Here, we propose data imputation and feature engineering methods that employ pre-trained large language models (LLMs) to enhance machine learning performance on scarce, heterogeneous datasets, demonstrated on graphene CVD synthesis data and the ML-HydPARK hydrogen storage dataset. GPT models perform data imputation via tailored prompting and semantic normalization of inconsistently reported features through embeddings, for example, to harmonize the complex nomenclature of CVD substrates. Beyond yielding more diverse and richer feature representations than traditional methods such as K-nearest neighbors (KNN) and Multivariate Imputation by Chained Equations (MICE), LLM-based data imputation is evaluated against dataset characteristics and prompting strategies. We vary the level of autonomy granted to the LLM, from generic prompting that leverages pre-trained knowledge for autonomous data generation to data-informed prompting that constrains outputs using target-specific information, and demonstrate which level of autonomy yields superior imputation performance across datasets and feature types. The proposed data engineering methods markedly improve downstream performance; for example, in graphene layer number classification using a support vector machine (SVM), binary accuracy increases from 39% to 65% and ternary accuracy from 52% to 72%. Fine-tuning experiments on both datasets show that combining our proposed LLM-based data imputation and feature encoding methods with numerical machine learning predictors outperforms standalone fine-tuned LLM predictors in data-scarce settings. The proposed strategies emphasize data enhancement techniques rather than refining learning architectures or regularizing loss functions, offering a broadly applicable framework for improving machine learning performance on scarce, inhomogeneous datasets.

Chemical vapor deposition↗

Integrated Hourly Meteorological Database of 20 Meteorological Stations (1981-2022) for Watershed Function SFA Hydrological Modeling

This dataset contains (a) a script “R_met_integrated_for_modeling.R”, and (b) associated input CSV files: 3 CSV files per location to create a 5-variable integrated meteorological dataset file (air temperature, precipitation, wind speed, relative humidity, and solar radiation) for 19 meteorological stations and 1 location within Trail Creek from the modeling team within the East River Community Observatory as part of the Watershed Function Scientific Focus Area (SFA). As meteorological forcings varied across the watershed, a high-frequency database is needed to ensure consistency in the data analysis and modeling. We evaluated several data sources, including gridded meteorological products and field data from meteorological stations. We determined that our modeling efforts required multiple data sources to meet all their needs. As output, this dataset contains (c) a single CSV data file (*_1981-2022.csv) for each location (20 CSV output files total) containing hourly time series data for 1981 to 2022 and (d) five PNG files of time series and density plots for each variable per location (100 PNG files). Detailed location metadata is contained within the Integrated_Met_Database_Locations.csv file for each point location included within this dataset, obtained from Varadharajan et al., 2023 doi:10.15485/1660962. This dataset also includes (e) a file-level metadata (flmd.csv) file that lists each file contained in the dataset with associated metadata and (f) a data dictionary (dd.csv) file that contains column/row headers used throughout the files along with a definition, units, and data type. Review the (g) ReadMe_Integrated_Met_Database.pdf file for additional details on the script, methods, and structure of the dataset.The script integrates Northwest Alliance for Computational Science and Engineering’s PRISM gridded data product, National Oceanic and Atmospheric Administration’s NCEP-NCAR Reanalysis 1 gridded data product (through the `RCNEP` R package, Kemp et al., doi:10.32614/CRAN.package.RNCEP), and analytical-based calculations. Further, this script downscales the input data into hourly frequency, which is necessary for the modeling efforts.

54 ENVIRONMENTAL SCIENCES↗

Aerobic respiration controls on shale weathering, Geochimica et Cosmochimica Acta, 2023: Dataset

This data package was generated in order to support the development of a deep-time weathering model and to assess the coupling between shale weathering and aerobic respiration in the paper “Aerobic respiration controls on shale weathering” by Stolze et al., Geochimica et Cosmochimica Acta (2023). The package contains two csv files providing the average CO2(g) concentration profiles [ppm] and mineral concentration profiles [wt%], respectively. The CO2(g) concentration profiles were measured in the vicinity of the monitoring well PLM2 between January 2018 and April 2019. The gas samples were collected in the unsaturated zone to a depth of 1.52 m. The mineral concentration profiles were determined by X-Ray diffraction (XRD). The XRD measurements were performed on sub-core samples collected in the monitoring well PLM3 down to a depth of 7.01 m. The dataset additionally includes a file-level metadata (flmd.csv) file that lists each file contained in the dataset with associated metadata; and a data dictionary (dd.csv) file that contains column/row headers used throughout the files along with a definition, units, and data type.Update on 2024-05-28: Revised versions of the CSV data files (CO2_data_GCA_Stolze_et_al_2023.csv and XRD_data_GCA_Stolze_et_al_2023.csv) were made to apply ESS-DIVE's CSV reporting format guidelines. Updated versions of the File Level Metadata (v2_20240528_flmd.csv) and Data Dictionary (v2_20240528_dd.csv) files were updated to reflect the changes made to the CSV files.

54 ENVIRONMENTAL SCIENCES↗

SG50 Data-format Requirement Document for an Automatically Readable, Comprehensive and Curated Experimental Reaction Database MEDUSA

This report constitutes the requirement document that guides the development of the experimental reaction database, MEDUSAL (Machine-readable Experimental Data User App & Library), created by OECD/NEA/WPEC SubGroup 50. Experimental reaction data are usually stored in the EXFOR library in EXFOR format. With MEDUSAL, the WPEC sub-group 50 wants to go beyond the EXFOR format and database to generate a library that is (a) automatically readable, (b) comprehensive, and (c) curated.

Nuclear Criticality Safety Program (NCSP)↗

COMPASS-FME Terrestrial Ecosystem Manipulation to Probe the Effects of Storm Treatments (TEMPEST) Experiment Level 1 Sensor Data v1-2

This is the version 1-2 Level 1 (L1) data release for COMPASS-FME environmental sensors located at our Terrestrial Ecosystem Manipulation to Probe the Effects of Storm Treatments (TEMPEST) experimental site. This manipulative, ecosystem-scale TEMPEST experiment addresses the potential for freshwater and estuarine-water disturbance events to alter tree function, species composition, and ecosystem processes in a deciduous coastal forest in MD, USA. The experiment uses a large-unit (2000 m2), un-replicated experimental design, with three 50 m × 40 m plots serving as control, freshwater, and estuarine-water treatments.L1 data are close to raw, but are units-transformed and have out-of-instrument-bounds and out-of-service flags added. Duplicates and missing data are removed but otherwise these data are not filtered, and have not been subject to any additional algorithmic or human QA/QC. Any scientific analyses of L1 data should be performed with care. **This dataset will be updated quarterly with new data for the duration of the project**This dataset includes:- An overall dataset README file that describes the current version, gives citation and contact information, etc.- Site- and year-specific folders, each holding up to 12 CSV (comma separated value) data files for each site and plot in that year.- Metadata files within each site-year folder provide full information on data units, expected ranges, contact information, detailed flood times, as well as a general description of the site.- Environmental sensor types that appear in the data files include weather (ClimaVUE50, CS, RM Young, and LI instruments in the graphs below); soil conditions (TEROS12); soil redox state (Redox); groundwater variables (AquaTROLL200 and AquaTROLL600); open water sondes (Exo); tree sap velocity (Sapflow); and system voltage and state (Datalogger). Data are normally logged every 15 minutes.Please see v1-2 TEMPEST L1 Sensor Package Quick Start.pdf for detailed information on data package structure, temporal coverage, and versioning.The TEMPEST flood events occurred on the following dates. They lasted for ~10 hours each day and delivered ~80,000 gallons to each plot; many data streams are available at 1 or 5 minute frequency during these periods.* Tests: Aug 25 (fresh plot) and Sep 9 (salt plot), 2021* TEMPEST 1: June 22, 2022* TEMPEST 2: June 6-7, 2023* TEMPEST 3: June 11-13, 2024

54 ENVIRONMENTAL SCIENCES↗

Model Data Archive Associated with Manuscript "Fire-altered Carbon Pools Create Disturbance Memory in Stream Dissolved Organic Carbon"

This data package supports the publication “Fire-altered Carbon Pools Create Disturbance Memory in Stream Dissolved Organic Carbon” by Li et al. (2026). The package contains processed model inputs, configuration files, restart files, simulation outputs, scripts, and visualization products used to evaluate post-fire dissolved organic carbon (DOC) dynamics in the Naches River Watershed, Washington, USA, following the 2021 Schneider Springs Fire. The modeling workflow couples ELM-BGC, the biogeochemistry-enabled Energy Exascale Earth System Model Land Model; ATS, the Advanced Terrestrial Simulator for integrated surface-subsurface hydrology; and PFLOTRAN, a reactive transport model for multicomponent aqueous geochemistry. Together, these models simulate how wildfire-induced changes in vegetation, litter, coarse woody debris, and soil organic matter influence DOC production, transport, and reaction from burned hillslopes to stream networks. The archive includes preprocessed meteorological, geospatial, hydrologic, and biogeochemical forcing data; ELM-BGC-derived DOC source terms; ATS mesh files; PFLOTRAN reactive-transport inputs; model configuration files; spin-up and transient restart files; watershed-scale diagnostic outputs; stream concentration time series; and figures or visualization files used to inspect and reproduce key results. File types include Hierarchical Data Format 5 (HDF5) files for gridded forcing and model-coupling data, model input and configuration files for ELM-BGC, ATS, and PFLOTRAN, restart and simulation-output files generated by the modeling workflow, tabular or time-series diagnostic outputs, scripts for post-processing and figure generation, and image or visualization products associated with the manuscript. Use of the package depends on the intended task. Re-running the simulations requires the relevant modeling software, including ELM-BGC, ATS, and PFLOTRAN as ATS's geochemical engine. Inspecting outputs and reproducing figures requires Python with scientific plotting libraries such as Matplotlib, and three-dimensional model outputs may be viewed with ParaView. Geographic information system files or maps may be inspected with ArcGIS Pro or comparable GIS software. The data package is intended to enable traceability, reuse, and partial reproduction of the coupled land-to-watershed hydro-biogeochemical modeling workflow used to test how wildfire disturbance affects terrestrial carbon pools and downstream DOC dynamics.

ATS↗