Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “data versions”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

IM3 Phase 2 Official Simulations: GCAM-Demeter-SELECT Annualized Land Use and Land Cover, Wood Harvest and Fertilization Data with Dynamic Urbanization Harmonized to CLM Land Definitions at 0.125 Degrees

Annualized land use land cover data, including wood harvest and fertilizer use data from the Global Change Analysis Model (GCAM) downscaled to 0.125 degrees for couping with the Community Land Model (CLM). GCAM here refers to GCAM-USA v5.3.im3 which has an enhanced electricity sector and an updated data system needed to represent regional to local scale dynamics. Data is also harmonized with future urbanization projections from the Spatially-Explicit, Long-term, Empirical City developmenT (SELECT) model. Projections/Data are generated using the demeter land use and land cover downscaling model. Original projections were generated at 0.05 degrees before being aggregared to 0.125 degrees. Projections are available for 8 alternative scenarios. Two versions of final data are included- one with managed forests or harvested forest area per pixel broken out and one with the same aggregated into total forests. Following folders are included: demeter_78_PFT_output:This is the final output of dynamic land use land cover change for 78 PFTs as required by CLM raw_outputs_incl_managed_forest: This is the final output but with managed forests broken out as a different PFT. Essentially a 79th PFT is added. wood_harvest_outputs: Wood harvest output per pixel in gC/m2 fertilization_outputs: Fertilizer use per pixel in gN/m2 Each NetCDF file in each folder represents a projection for a separate year, scenario. Land use outputs are organized as PFT level data saved as subdata. Link to GCAM version used- https://data.msdlive.org/records/yb23g-44274 Link to SELECT documentation -https://www.sciencedirect.com/science/article/pii/S1364815219301707 Link to CLM documentation- - https://www.cesm.ucar.edu/models/clm In case of questions contact- kanishka.narayan@pnnl.gov

GCAM-USA↗

Evaluation of Sandia NCS Benchmark Suite Updates

The Sandia Nuclear Criticality Safety (NCS) program’s benchmark suite was recently updated. This suite is used to ensure that NCS calculations using computer-based neutron transportation codes have an established baseline comparison of calculated versus known experimental results. The Evaluated Nuclear Data File (ENDF) version used for the MCNP models in the suite was changed from ENDF/B-VII.1 to ENDF/B-VIII.0. Additionally, relevant thermal scattering law data libraries (TSLs) were updated. The sensitivity of the calculational bias of each benchmark model to these changes is discussed. Implementation of the ENDF/B-VIII.0 library and updated TSLs results in improvements to bias distribution in the intermediate enriched uranium, plutonium, and mixed uranium–plutonium (IEU, PU, and MIX) fissionable material benchmark categories, but a small bias increase in low- and high-enriched uranium categories (LEU and HEU, respectively). The results also highlight the sensitivity of the benchmarks, with average lethargy of neutrons causing fission energies (EALF) in the intermediate energy range to ENDF/B library changes. The most numerous bias changes were observed in the thermal energy region when transitioning from ENDF/B-VII.1 to ENDF/B-VIII.0. In conclusion, most of the unique bias changes observed in MCNP 6.3.0 between the two nuclear data libraries were in the LEU-COMP-THERM evaluation subset.

ICSBEP↗

Interaction of SPI pellets with plasma on JET and associated disruptions

Abstract The presented data refer to the Shattered Pellet Injector (SPI) experiments carried out at JET in 2019–2020. This paper is a full journal version of the data originally presented as posters at TMPDM_2020 and EPS_2021. This paper presents various aspects of the interaction of pellets with plasma and associated disruptions. The experiment was performed with I p = (1.1–3.1) MA plasmas and mainly with Ne + D 2 pellet composition, but also with Ar pellets. The Current Quench (CQ) time, τ 80−20 , is the key characteristic of mitigation effectiveness. A pellet with a high content of Ne or Ar can reduce the CQ duration below the upper required JET threshold. Plasmas with high (thermal + internal poloidal magnetic) pre-disruptive plasma energy require a high content of Ne pellets to obtain a short CQ duration. Pellets with a small amount of Ne (and accordingly large amount of D), instead of causing a mitigated CQ, create the conditions for a ‘cold’ Vertical Displacement Events (VDE). The SPI was applied to plasma with different status: mainly to normal (‘healthy’) plasma, i.e. not prone to disruption, post-disruptive and VDE plasma. This study shows that SPI effectiveness in terms of CQ duration and, accordingly, EM loads does not depend on the state of the plasma, whether it is ‘healthy’ or post-disruptive plasma. SPI has been shown to reduce the axisymmetric vertical vessel reaction forces by about (30–40) % compared to unmitigated disruptions. On JET, the VDE, whether ‘hot’ or ‘cold’, always creates the conditions for a toroidal asymmetry in the plasma, so the VDE on the JET is referred to as Asymmetric VDE (AVDE). The interrupting of VDE and prevention of AVDE with SPI has been demonstrated. Thus, the effectiveness of disruption mitigation using SPI has been confirmed.

Gerasimov, S. N. (ORCID:0009000237937211)↗

nys_psy (NYgrid Model Translation to the Sienna framework) (SWR-25-63)

This repository contains the translation of the NYgrid model, developed by the Anderson Energy Lab at Cornell University, into the Sienna Framework. The baseline model is based on 2019 data. The 2040 version of the model features a unified, correlated dataset of various generation and load profiles spanning 22 years. The methodology used to generate these data is detailed in the following paper The scripts for data generation are available in the ny-clcpa2050 repository. Note: While this test system is designed to simulate the power flow of the New York State transmission system, it does not represent the actual transmission network.

Liu, Vivienne [National Renewable Energy Laborator↗

PLUSWIND Hourly Plant-Level Power

These data contain the hourly plant-level power time series from PLUSWIND. The time series is derived from the HRRR model and corrected for density and losses. The PLUSWIND dataset hosted on the Wind Data Hub here (https://a2e.energy.gov/project/pluswind) contains additional versions with meteorological corrections. Data are provided for years 2018 - 2021. The data hosted at this location include conversion from capacity factor output (normalized output) to total hourly production (units of MW).

17 WIND ENERGY↗

Sandia National Laboratories Ecosystem for Open Science: Metadata Schema v0.2 Description.

The Ecosystem for Open Science (eOS) initiative was established in 2019. Its objective is improving openness and sharing of data and information across Defense Nuclear Nonproliferation (DNN) Research and Development (R&D) activities. To support this initiative, the eOS team at Sandia National Laboratories (SNL) developed metadata and data standards and proposed a machine-readable metadata schema. The nuclear explosion monitoring field was selected as a focus area due to its the wide range of pertinent phenomenologies.We developed the DCAT-eOS-AP metadata schema extending the Data Catalog Vocabulary version 2 (DCATv2) standard using an application profile (AP), to fit the needs of multi-disciplinary NA-22 projects. The DCAT-eOS-AP metadata schema describes data at different levels of granularity ranging from general descriptions to more domain-specific granular metadata. Its implementation and serialization is flexible with the ability to include new file or data types. Thus, it will scale with the ever-increasing data management needs of government research. Due to the multitude of phenomenologies represented in the DCAT-eOS-AP schema, we anticipate that it will be easily extensible to various projects across many DOE mission areas. This document describes data management challenges faced within the DNN R&D portfolio and provides insight on how metadata and data standards/guidelines combined with a comprehensive metadata schema can add value to programs throughout the Department of Energy (DOE). It reviews the importance of metadata standards, FAIR (Findability, Accessibility, Interoperability, and Reusability) data principles, and metadata schemas. Additionally, it summarizes input from subject matter experts (SME) at SNL and other National Laboratories that resulted in metadata and data standards/guidelines encompassing domains relevant to NA-22 projects. Finally, we discuss the DCAT-eOS-AP metadata development. Implementation recommendations and future development directions are included for those keen on adopting the DCAT-eOS-AP metadata schema.

96 KNOWLEDGE MANAGEMENT AND PRESERVATION↗

Data from: "Warming and provenance limit tree recruitment across and beyond the elevation range of subalpine forest"

This data package contains data used to support conclusions drawn in “Warming and provenance limit tree recruitment across and beyond the elevation range of subalpine forest”, by Kueppers et al. 2017. Data were collected in field sites within the Alpine Treeline Warming Experiment (ATWE), located on Niwot Ridge, on the eastern slope of the Colorado Rocky Mountains, USA. Files containing geospatial data are also included, to provide additional locational context.There are four document formats associated with this archive: three comma-separated values (.csv) files, three Microsoft Excel (.xlsx) files, one .pdf data user’s guide, four keyhole markup language (.kml) files, and a compressed folder containing seven ESRI shapefiles (.shp). The .csv files can be opened using any simple text-editor software, R, or Microsoft Excel. The .xlsx files can only be opened using Microsoft Excel. The .kml file can be opened by Google Earth and Google Maps, and the shapefiles can be opened with any GIS application compatible with the file type, such as ESRI’s ArcGIS, and QGIS.We provide two versions of the seedling data file: “PIEN_PIFLseedlings20150522_20150525rev12222020.csv/.xlsx” (hereafter PIEN_PIFLseedlings2015) and “PIEN_PIFLseedlings20160408rev12222020.csv/.xlsx” (hereafter PIEN_PIFLseedlings2016). PIEN_PIFLseedlings2015 contains the data we used in the paper. PIEN_PIFLseedlings2016 contains an updated version of these data that includes sampling from later years. The main differences between the two files lie in the columns titled “k[YEAR],” which describe the number of seedlings that were killed in a particular year. In PIEN_PIFLseedlings2016, there also is an additional year of data for k2015, and k2014 also has additional data input for the 2014 cohort. Additionally, in years 2010-2014, there are minor differences in the number of seedlings killed -- in as few as 0 plots (in 2011) to as many as 5 plots (in 2014) -- due to errors in data input that were rectified in later years.------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------Upslope range shifts by subalpine tree species are a widely anticipated effect of climate change. Climate niche models predict subalpine forests to expand upslope, given more suitable growing conditions for adult trees. However, these models do not take into account climates required for successful seedling recruitment and establishment, an essential element for expansion. Further, localized upper treeline populations are hypothesized to contain favorable traits for colonizing the alpine. To test these expectations and to expand our knowledge of seedling recruitment under climate change, we designed a common garden, climate-warming experiment spread across an elevation gradient at Niwot Ridge in the Colorado Rocky Mountains. We focus on two widespread Western North American species, Engelmann spruce (Picea engelmannii Parry ex. Engelm) and limber pine (Pinus flexilis James), which occur at treeline. While the former is considered a late-seral species more tolerant of shade, limber pine is a shade-intolerant pioneer species able to establish on infertile sites.Every autumn, seeds of the two species were collected from high- (3370 m–3570 m) and low-provenance (2910–3240 m) sources close to the experimental sites and sown in our plots. A subset of plots were heated and another subset watered over the summer months to offset the effects of warming. Across five years, we found that seeds originating from low elevation recruited more strongly for both species, although this provenance difference diminished by the fourth year for Engelmann spruce, likely due to small sample sizes. Despite the recruitment of low-provenance seed, warming treatments decreased recruitment at all elevations. Combining this with the likeliness and availability of lower-quality, high provenance seed moving upslope at the treeline, tree migration into the alpine may be slowed. Overall, our findings suggest that the hardier limber pine is likely to become a more significant species in subalpine forest communities in the future, while the more sensitive Engelmann spruce may experience range contraction.

54 ENVIRONMENTAL SCIENCES↗

Pyrolysis Modeling of PMMA decomposition studied by TGA

Data from four TGA experiments conducted at Sandia National Laboratories was used for determination of a pyrolysis model using a commercial thermokinetics program developed by Netzsch Instruments (Kinetics NEO, version 2.1). The data measured at 1 K/min and the average of three measurements at 50 K/min were used as input into Kinetics NEO. The model was developed using data in the range 373 to 773 K. An initial estimate of the energy of activation (E) and pre-exponential constant (A) were determined from the model-free Friedman approach.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Empirical Validation of UBEM: An Assessment of Bias in Urban Building Energy Modeling for Chicago

Residential and commercial buildings currently account for 30% of total global final energy consumption. Urban-scale building energy modeling (UBEM) can enable scalable investments and unlock building improvements by quantifying energy, demand, emissions, and cost reductions of specific measures or packages for building-specific technologies in large geographic regions. While the sophistication of UBEM data sources and technologies have increased dramatically in the past decade, there remains a knowledge gap for empirical validation and sources of bias between building-specific energy models and measured data at varying geographic scales.As UBEM continues to develop, systemic analysis of accuracy, bias, and limitations of the resulting models is necessary to inform best practices and move toward standardization. These are characterized for the Automatic Building Energy Modeling (AutoBEM) software suite with an initial case study involving metered electricity consumption data from 247,188 buildings in Chicago, Illinois, USA - averaged across years 2019-2021 - compared to the following datasets: (1) the AutoBEM-generated nation-scale Model America version 2 (MAv2) data for 596,064 buildings, (2) tax assessor data for 579,829 buildings, (3) tax assessor data filled with MAv2, and (4) 102 representative dynamic archetypes. The accuracy is reported for every building type and vintage combination, along with multiple sources of bias for unique building descriptors. The AutoBEM simulation workflow produced energy consumption estimates that closely match aggregated metered electricity consumption data for different types of buildings constructed during various time periods at the city scale - with initial normalized mean bias error of 10.9%, and 1.1% after removing outliers. Contribution of statistically significant factors including building type, land use, age, and size to variance in UBEM bias is quantified.

Garg, Ankur↗

Sensitivity of Regional WRF‐Chem Air Quality and Weather Simulations to Biomass‐Burning Emission Data Sets: A Case Study of the Impact of Canadian Wildfire on the US°

This study focuses on the period from June 26 to 29, 2023, when record‐breaking Canadian wildfires severely impacted air quality in the Midwest United States. Using the Weather Research and Forecasting Model with Chemistry (WRF‐Chem) and four biomass‐burning data sets (Fire Inventory from NCAR version 1, Fire Inventory from NCAR version 2.5, Quick Fire Emissions Data set [QFED], and Regional ABI‐VIIRS Emission), we analyzed aerosol transport from Canada to the US and assessed the model's accuracy in predicting PM 2.5 , O 3 , CO and aerosol weather feedback. Model simulations were compared with ground‐based and remote sensing observations as well as field measurements from the Community Research on Climate and Urban Science (CROCUS) project. Our findings show that the movement of a low‐pressure system from the Great Lakes to the Atlantic, combined with the high‐pressure system over the Atlantic, caused the transport of aerosols from Canadian wildfires to the US. Results show WRF‐Chem significantly underestimated key atmospheric components: aerosol optical depth (AOD) by over 50%, PM 2.5 by 65%–90% and peak O 3 concentrations by 50%–55% across four biomass burning data sets. Additionally, CO and NO 2 concentrations were underpredicted. The substantial underestimation of PM 2.5 led to an overestimation of temperature by up to 3.6 °C primarily due to excessive downward shortwave radiation, which resulted from the underestimation of direct aerosol effects and an increase in sensible heat flux. Among the biomass‐burning data sets, QFED produced the most accurate AOD and PM 2.5 predictions due to improved wildfire emission estimates, leading to a 1.0 to 1.5 °C reduction in temperature overestimation during the daytime. These findings underscore the need for improving wildfire emission estimates for trace gases and aerosols to enhance air quality and weather feedback predictions.

WRF-chem model↗

Solute chemistry for streams draining geomorphic features and varying land cover gradients in the East River watershed, Colorado

This dataset contains solute chemistry data and GPS coordinates for surface water samples collected in the East River watershed in the Elk and West Elk Ranges of Colorado from August 3rd through August 11th, 2022. Samples were collected from streams and groundwater seeps as part of a comprehensive analysis within the watershed of how different geomorphic features (such as landslides and rock glaciers) along with variations in land cover (such as the presence of vegetation) affect stream water quality in this Rocky Mountain headwater environment. This dataset includes concentrations of the following in a single csv file: fluoride, chloride, sulfate, nitrate, calcium, potassium, magnesium, sodium, silicon, and total organic carbon (TOC). All concentrations are reported in mg/L. The sulfate data in this version have been updated to account for an earlier ion chromatograph instrument method error whereby sulfate and bromide peaks overlapped in calibration solution samples, biasing calculated sulfate concentrations low by 30-50%. Sulfate data reported herein are from a re-run of the samples performed in December 2024 using a corrected instrument method. All samples were stored under refrigerated conditions and filtered on initial collection.

54 ENVIRONMENTAL SCIENCES↗

The Fourth Catalog of Active Galactic Nuclei Detected by the Fermi Large Area Telescope: Data Release 3

Abstract An incremental version of the fourth catalog of active galactic nuclei (AGNs) detected by the Fermi Large Area Telescope is presented. This version (4LAC-DR3) derives from the third data release of the 4FGL catalog based on 12 yr of E > 50 MeV gamma-ray data, where the spectral parameters, spectral energy distributions (SEDs), yearly light curves, and associations have been updated for all sources. The new reported AGNs include 587 blazar candidates and four radio galaxies. We describe the properties of the new sample and outline changes affecting the previously published one. We also introduce two new parameters in this release, namely the peak energy of the SED high-energy component and the corresponding flux. These parameters allow an assessment of the Compton dominance, the ratio of the inverse-Compton to the synchrotron-peak luminosities, without relying on X-ray data.

79 ASTRONOMY AND ASTROPHYSICS↗

The ENDF/B Nuclear Data Library and Its Impact on Reactor Simulations

The ENDF/B library, which is developed, maintained, and distributed by the Cross Section Evaluation Working Group, is the main source of nuclear data for analyses and computational simulations in nuclear applications, like different nuclear reactors concepts, radiation shielding, medical applications, astrophysics, etc. The library is constantly being improved and updated, with each release bringing an optimal representation of nuclear interactions as they are understood in their time. The most recent release, ENDF/B-VIII.1, represents a significant improvement in terms of the performance and consistency of the measured differential data relative to previous versions, as it combines the most recent experimental differential data and advanced theoretical nuclear models. As one of the many highlights, ENDF/B-VIII.1 restores a high-burnup depletion performance, comparable to ENDF/B-VII.1, that had been degraded in ENDF/B-VIII.0, while further improving the performance in criticality benchmarks, such as those in the Mosteller’s suite. Additionally noteworthy is the improved performance of ENDF/B-VIII.1 in radiation shielding and thick-target leakage spectrum integral experiments, which are also important for fusion and reactor applications. In this work we present in a very summarized way the main updates implemented in the ENDF/B-VIII.1 release and its main impacts, and also begin to delineate the path forward as to what to expect in the future for the next ENDF/B release, which, based on the timeline of the past few releases, is estimated to happen around 5 years from now. We emphasize that the ENDF/B-VIII.1 release was the product of an enormous collaborative effort among many authors and that for a complete detailed picture, the reader is strongly encouraged to refer to the article accompanying the release, which is currently in the publication process, but available as preprint [G. P. A. Nobre et al. arXiv Preprint arXiv: 2511.03564 (2025)].

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Hyper Suprime-Cam Legacy Archive

Abstract We present the launch of the Hyper Suprime-Cam Legacy Archive (HSCLA), a public archive of processed, science-ready data from the Hyper Suprime-Cam (HSC). HSC is an optical wide-field imager installed at the prime focus of the Subaru Telescope and has been in operation since 2014. While ∼1/3 of the total observing time of HSC has been used for the Subaru Strategic Program (SSP), the remainder of the time is used for Principal Investigator (PI)-based programs. We have processed the data from these PI-based programs and make the processed, high-quality data available to the community through HSCLA. The current version of HSCLA includes data taken in the first year of science operation, 2014. We provide both individual and coadd images as well as photometric catalogs. The photometric catalog from the coadd is loaded to the database, which offers a fast access to the large catalog. There are other online tools such as an image browser and an image cutout tool and they will be useful for science analyses. The coadd images reach 24–27th magnitudes at 5σ for point sources and cover approximately 580 square degrees in at least one filter with 150 million objects in total. We perform extensive quality assurance tests and verify that the photometric and astrometric quality of the data is good enough for most scientific explorations. However, the data are not without problems and users are referred to the list of known issues before exploiting the data for science. All the data and documentations can be found at the data release site, 〈https://hscla.mtk.nao.ac.jp/〉.

Tanaka, Masayuki↗

High-Resolution Fire Weather Index Data for the Conterminous US (1980–2099), Version 1

This dataset presents a suite of high-resolution fire weather index datasets calculated from observation (gridMet, Livneh, Daymet V4), reanalysis (AgERA5), downscaled hydro-climate projections over the conterminous United States (CONUS) based on multiple selected Global Climate Models (GCMs) from the Coupled Model Intercomparison Project Phase 6 (CMIP6). Aside from the daily FWI datasets, we also include a set of FWI extreme indicators at annual, seasonal, and monthly scales, including 1) fwixx: maximum FWI; 2) fwisa: mean FWI. For annual fwisa, it refers to the season with the maximum seasonal average; 3) fwils: Length of fire season over a specified period, where fire season is defined as the days exceeding the median value of the normalized FWI during the reference period (1980-1984); 4) fwixd: Number of extreme fire weather days over a specified period, where extreme day is defined as the day with FWI > the 95th percentile of the FWI during the reference period (1980-1984). All FWI datasets cover 1980-2020 baseline and the model simulated products including the downscaled products additionally include 2021-2099 near-future periods under the high-end (SSP585) emission scenario.

54 ENVIRONMENTAL SCIENCES↗

Data and scripts associated with “Moisture content modulates DOM thermodynamic regulation of oxygen consumption in drying streambed sediments”

This data package is associated with the publication “Moisture content modulates DOM thermodynamic regulation of oxygen consumption in drying streambed sediments” published in Scientific Reports (Garayburu-Caruso et al., 2026). The package contains processed data products and scripts used to quantify how drying and re-inundation of riverbed sediments influence dissolved organic matter (DOM) thermodynamic properties and their relationship with sediment oxygen (O₂) consumption across 33 stream sites in the contiguous United States. The data package contains DOM thermodynamic metrics (e.g., Gibbs free energy of carbon oxidation and thermodynamic efficiency), and O₂ consumption along with watershed-scale climate and land-cover metrics used as explanatory variables in the analyses. Underlying unprocessed and processed ultrahigh-resolution mass spectrometry data, oxygen consumption rates from laboratory moisture-manipulation experiments, within-sample environmental properties, sediment moisture content and contextual field measurements are archived separately at https://data.ess-dive.lbl.gov/datasets/doi:10.15485/2428003 (Laan et al., 2024) and https://data.ess-dive.lbl.gov/datasets/doi:10.15485/1923689 (Forbes et al.,2023). A preliminary version of this data package was published in February 2026 at the time of manuscript submission. It was updated in June 2026, at the time of manuscript acceptance, to include the finalized data and additional metadata (readme, data dictionary, and file level metadata). For details on how to navigate data packages generated by this project, see https://data.ess-dive.lbl.gov/portals/PNNLRiverCorridorSFA/About. In addition to a readme, this data package also includes a file-level metadata (FLMD) file that describes each file and a data dictionary (DD) that describes all column/row headers and variable definitions. At the top level, the data package is organized into five main folders: (1) Data, (2)Figures, (3) Map, (4) GAM_Reulsts, and (5) src. The Data folder contains analysis-ready tabular files with oxygen consumption rates, DOM thermodynamic properties by site and treatment, site-level environmental variables, watershed-scale metrics, and other derived variables referenced in the manuscript. The Figures folder contains static image files associated with the main text and supplemental figures, while the Map folder includes spatial data and map-layer files used to create the sampling-location map. The GAM results folder contains the results for each of the general additive model (GAM).The src folder contains R scripts used to perform data processing, statistical analyses (including clustering, generalized additive models, and threshold analysis), and figure generation. This data package is associated with a GitHub repository found at https://github.com/WHONDRS-Hub/ECA_DOM_Thermodynamics.

Dissolved organic matter↗

WHONDRS River Corridor Sediment and Water Geochemistry and In Situ Sensor Data from 7 Perennial and 7 Intermittent Streams across San Antonio, Texas (v3)

This dataset supports a broader study examining the effects of intermittency on sediment respiration. The dataset provides sediment and surface water geochemistry and in situ sensor data from 7 perennial and 7 intermittent streams in San Antonio, Texas. Each stream/site was visited both in summer during base flow (July-September 2023) and winter during peak flow (January-February 2024). Related data were collected and will be published separately in collaboration with A. Veach. The data package was originally published in April 2025. It was updated in June 2025 (v2; modified and new files) and September 2025 (v3; modified files). See the change history section in the readme for more details. For details on how to navigate data packages generated by this project, see https://data.ess-dive.lbl.gov/portals/PNNLRiverCorridorSFA/About. This dataset is comprised of two folders of field photos and videos, one folder of raw Fourier transform ion cyclotron resonance mass spectrometry (FTICR-MS) data and one main data folder containing (1) file-level metadata; (2) data dictionary; (3) field metadata; (4) readme; (5) international generic sample number (IGSN) mapping file; (6) field protocol; (7) a subfolder with sample data; and (8) a subfolder with sensor data. The sample data subfolder contains (1) surface water and sediment dissolved organic carbon (DOC, measured as non-purgeable organic carbon, NPOC) data and averages; (2) surface water and sediment total nitrogen data and averages; (3) sediment grain size data; (4) sediment iron (II) data and averages; (5) wet sediment mass, dry sediment mass, water mass, and wet sediment volume in incubation and sediment ICR vials; (7) sediment incubation respiration rate data and averages; (8) normalized respiration rate data and averages; (9) methods codes; (10) sediment percent carbon and nitrogen; (11) sediment X-ray diffraction (XRD) data; (12) gravimetric moisture and averages; (13) a subfolder with sediment incubation respiration data, scripts, and plots; (14) surface water and sediment FTICR methods; and (15) a subfolder of 9.4 Tesla (9.4T) FTICR-MS data. This folder contains five subfolders, one containing the sediment .xml data files, one containing the water .xml files, one containing the sediment CoreMS output files, one containing the water CoreMS output files, and the other containing instructions and scripts for processing the files in CoreMS (https://github.com/EMSL-Computing/CoreMS). The sensor data subfolder contains (1) a subfolder with miniDOT dissolved oxygen and temperature data and plots; (2) miniDOT dissolved oxygen and temperature summary data; and (3) miniDOT installation methods. All files are .csv, .pdf, .R, .xml, .d, .html, .Rmd, .py, .cal, .json, .jpg, .jpeg, .png, .mov, or .mp4. CORRECTION: The data processing methods for FTICR described in “v3_WHONDRS_AV1_Methods_Codes.csv” mistakenly indicate that users should process the data in Formultitude. The corrected description should read: “Both unprocessed and processed data are provided to allow users flexibility in data processing. Instructions and scripts for processing the data using CoreMS are included.” CORRECTION: Carbon and nitrogen content are reported as percentages. The current column headers "01395_C_percent_per_mg" and "01397_N_percent_per_mg" are incorrect. These should read "01395_C_percent" and "01397_N_percent" and will be corrected in the next version of this data package.

54 ENVIRONMENTAL SCIENCES↗

KCG Baseline Air Pollution Station LIFSO2 v1 Data from December 2024 to March 2025

Sulfur Dioxide is a key precursor to the formation of new particles within the marine environment, yet commercial instrumentation lack sufficient precision or sensitivity to resolve the levels present in these environments. As such, the LIFSO2, a custom built fibre laser from the University of York, UK, was developed (based on the one developed by Rollins et al 2016). We are able to resolve down to the ppt level with this instrument. We present version 1 (v1) data for Dec 2024 to Mar 2025 for SO2 (1min time average) from the Cape-k precursors field campaign.

Sulfur dioxide (SO2) mixing ratio↗