Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “open data format”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 145 records · Page 8

Multimodal super-resolution: discovering hidden physics and its application to fusion plasmas

Understanding complex physical systems often requires integrating data from multiple diagnostics, each with limited resolution or coverage. We present a machine learning framework that reconstructs synthetic high-temporal-resolution data for a target diagnostic using information from other diagnostics, without direct target measurements during the inference. This multimodal super-resolution technique improves diagnostic robustness and enables monitoring even in case of measurement failures or degradation. Applied to fusion plasmas, our method targets edge-localized modes (ELMs), which can damage plasma-facing materials. By reconstructing super-resolution Thomson Scattering data from complementary diagnostics, we uncover fine-scale plasma dynamics and validate the role of resonant magnetic perturbations (RMPs) in ELM suppression through magnetic island formation. The approach provides new observation supporting the plasma profile flattening due to these islands. Our results demonstrate the framework’s ability to generate high-fidelity synthetic diagnostics, offering a powerful tool for ELM control development in future reactors like ITER. The approach is broadly transferable to other domains facing sparse, incomplete, or degraded diagnostic data, opening new avenues for discovery.

Jalalvand, Azarakhsh [Princeton Univ., NJ (United ↗

An idea to explore: How an interdisciplinary undergraduate course exploring a global health challenge in molecular detail enabled science communication and collaboration in diverse audiences

Abstract Communication and collaboration are key science competencies that support sharing of scientific knowledge with experts and non‐experts alike. On the one hand, they facilitate interdisciplinary conversations between students, educators, and researchers, while on the other they improve public awareness, enable informed choices, and impact policy decisions. Herein, we describe an interdisciplinary undergraduate course focused on using data from various bioinformatics data resources to explore the molecular underpinnings of diabetes mellitus (Types 1 and 2) and introducing students to science communication. Building on course materials and original student‐generated artifacts, a series of collaborative activities engaged students, educators, researchers, healthcare professionals and community members in exploring, learning about, and discussing the molecular bases of diabetes. These collaborations generated novel educational materials and approaches to learning and presenting complex ideas about major global health challenges in formats accessible to diverse audiences.

59 BASIC BIOLOGICAL SCIENCES↗

Industry Facing PV Degradation Prediction Tool and Database to Enable a 50-Year Life Module

The Photovoltaic (PV) industry constantly aims for lower costs, higher-efficiency cells, and improved module designs. These trends lead to using new materials, designs, and manufacturing processes, resulting in a continually changing technological landscape. These changes can potentially introduce new, unknown degradation mechanisms and failure modes that are difficult to diagnose, analyze, test, and model. This introduces uncertainty into the expected lifetime of PV modules of 25 years. Furthermore, research efforts aim for up to 50 years of service life while keeping performance degradation at a minimum for decades of outdoor weathering - putting additional pressure on improving the accuracy of long-term durability and reliability assessments. Here there is a need to organize the existing degradation data into an accessible format and to provide industry relevant tools for extrapolation from laboratory to field conditions. Because the core of this type of analysis involves calculations that are complicated but ubiquitous for many degradation processes, an enhanced predictive modeling framework will facilitate the analysis to help researchers keep up with the rapid pace of technological changes. In this work, we present an online tool that can be used to search for and analyze degradation information and extrapolate PV module performance and durability to field exposure. A graphical user interface will aid in the understanding of the results. The prediction tool will be built modular and published as open source, enabling users to expand on the existing framework. We use an integration pipeline approach that allows us to leverage weather data from the National Solar Radiation Database to perform geospatial degradation analysis in the US and worldwide. Our repository will contain various degradation models and material parameters suitable for the reliability and durability assessment of materials and components deployed outdoors. We hope to become a repository that can be used for weathering and degradation analysis for various applications beyond the PV industry.

degradation↗

Binding, Release and Functionalization of Intact Pnictogen Tetrahedra Coordinated to Dicopper Complexes

The bridging MeCN ligand in the dicopper(I) complexes [(DPFN)Cu 2 (μ,η 1 : η 1 -MeCN)][X] 2 (X=weakly coordinating anion, NTf 2 (1 a), FAl[OC 6 F 10 (C 6 F 5 )] 3 (1 b), Al[OC(CF 3 ) 3 ] 4 (1 c)) was replaced by white phosphorus (P 4 ) or yellow arsenic (As 4 ) to yield [(DPFN)Cu 2 (μ,η 2 : η 2 -E 4 )][X] 2 (E=P (2 a-c), As (3 a-c)). The molecular structures in the solid state reveal novel coordination modes for E 4 tetrahedra bonded to coinage metal ions. Experimental data and quantum chemical computations provide information concerning perturbations to the bonding in coordinated E 4 tetrahedra. Reactions with N-heterocyclic carbenes (NHCs) led to replacement of the E 4 tetrahedra with release of P 4 or As 4 and formation of [(DPFN)Cu 2 (μ,η 1 : η 1 - Me NHC)][X] 2 (4 a,b) or to an opening of one E-E bond leading to an unusual E 4 butterfly structural motif in [(DPFN)Cu 2 (μ,η 1 : η 1 -E 4 Dipp NHC)][X] 2 (E=P (5 a,b), E=As (6)). With a cyclic alkyl amino carbene ( Et CAAC), cleavage of two As-As bonds was observed to give two isomers of [(DPFN)Cu 2 (μ,η 2 : η 2 -As 4 Et CAAC)][X] 2 (7 a,b) with an unusual As 4 -triangle+1 unit.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

The ePIC Simulation Campaign Workflow on the Open Science Grid

The ePIC collaboration is realizing the first experiment of the future Electron-Ion Collider (EIC) at the Brookhaven National Laboratory that will allow for a precision study of the nucleons and the nucleus at the scale of sea quarks and gluons through the study of electron-proton/ion collisions. This paper will discuss the current workflow for running centralized simulation campaigns for ePIC on the Open Science Grid (OSG) infrastructure. This involves monthly releases of ePIC software and container deployments to CVMFS, generation of input datasets in HepMC format according to collaboration-defined policy, using Snakemake in CI/CD for validation and benchmarking, and submitting jobs to the OSG condor scheduler for opportunistic running on available resources. File transfers utilize XrootD, and Rucio is used for data management. The workflow is continuously refined to improve daily throughput (currently 50-100k core hours per day) and minimize job failures. Since May 2023, monthly simulation campaigns employing the workflow have cumulatively used over 20 million core hours on the OSG and produced over 350 TB of simulation data. The campaigns incorporate simulations for the broad science program of the EIC and are actively used for the detector and physics studies in preparation of the Technical Design Report (TDR).

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

IM3 Projected US Data Center Locations

IM3 Projected US Data Center Locations This dataset contains model projections of new data center facilities in the contiguous United States (CONUS) through 2035 using the CERF – Data Centers model. Data center locations are modeled across four data center electricity demand growth scenarios (low, moderate, high, higher) and five market gravity scenarios (0%, 25%, 50%, 75%, 100%). Projected locations are intended to be regional representations of feasible siting locations in the future to assess potential grid and water stress impacts. The data center load growth scenarios correspond with the rates outlined in EPRI (2024) and include 3.71%, 5%, 10%, and 15% annual growth of electricity demand for data centers from 2023 values in 37 states across the CONUS. Market gravity scenarios correspond to the relative importance of proximity to data center markets or high population areas compared to locational cost in the siting algorithm. 0% market gravity means that siting decisions were entirely determined by the locational cost in each feasible location. 100% market gravity means that only market proximity was considered when siting. Other scenarios have weight placed on both components where total weight always equals 100%. Locational cost is dependent on facility cooling type and corresponding electricity cost, taxes, and other factors. Facility cooling type is spatially determined where high water stress and/or areas with high summer wet bulb temperatures are assumed to operate with mechanical cooling for a higher fraction of the year rather than evaporative cooling. Feasible data center siting areas are based on geospatial suitability raster data developed with open-source information. The following areas are excluded from siting: Areas within 300 m of a federal airport runway Waterbodies Areas with slope >16% Areas susceptible to sinkholes High coastal or inland flood risk areas Local, state, and federal parks, leisure areas, and cemeteries Areas >2 km away from electric substations Areas >5 km away from a municipal water supplier service area Areas >2 km away from high-speed fiber provider service territory Protected Areas Database of the United States (PAD-US) areas Railroads, major roadways, and minor roadways Military areas and training grounds NLCD developed lands Areas >0.8 km (0.5 miles) from NLCD developed lands Because we use open-source information, proprietary information that can influence siting decisions such as individual tax agreements with cities, detailed fiber line connectivity, electric grid power capacity agreements, and others, are not currently accounted for in the modeling process. Using specific building locations and footprints in the dataset for local planning purposes is not advised. Technical Information Geospatial data is provided in geojson format using the Albers Equal Area Conic (ESRI:102003) coordinate reference system. The datasets contain the following parameters: id - unique identification number within given scenario file growth_scenario – data center demand growth scenario market_gravity_weight – market gravity weight scenario (%) region – name of region (i.e., US State) total_cost_million_usd – locational siting cost ($million) campus_size_square_ft – total land acquired for data center facility (square ft) data_center_it_power_mw – IT power of data center facility (MW) mechanical_cooling_frac – fraction of year when data center uses mechanical cooling system water_cooling_frac– fraction of year when data center uses evaporative cooling system cooling_energy_demand_mwh – total annual facility energy demand for cooling (MWh) cooling_water_demand_mgy – total annual facility water demand for cooling (MG) cooling_water_consumption_mgy – total annual facility water consumed (MG) normalized_locational_cost – normalized total locational cost score for location normalized_gravity_score – normalized market gravity score for location weighted_siting_score – total weighted siting score of locational cost and gravity score geometry – polygon geometry of facility Acknowledgment IM3 is a multi-institutional effort led by Pacific Northwest National Laboratory and supported by the U.S. Department of Energy's Office of Science as part of research in MultiSector Dynamics, Earth and Environmental Systems Modeling Program. License This data is made available under a CCBY4.0 License Disclaimer This material was prepared as an account of work sponsored by an agency of the United States Government. Neither the United States Government nor the United States Department of Energy, nor the Contractor, nor any or their employees, nor any jurisdiction or organization that has cooperated in the development of these materials, makes any warranty, express or implied, or assumes any legal liability or responsibility for the accuracy, completeness, or usefulness or any information, apparatus, product, software, or process disclosed, or represents that its use would not infringe privately owned rights. Reference herein to any specific commercial product, process, or service by trade name, trademark, manufacturer, or otherwise does not necessarily constitute or imply its endorsement, recommendation, or favoring by the United States Government or any agency thereof, or Battelle Memorial Institute. The views and opinions of authors expressed herein do not necessarily state or reflect those of the United States Government or any agency thereof. PACIFIC NORTHWEST NATIONAL LABORATORYoperated byBATTELLEfor theUNITED STATES DEPARTMENT OF ENERGYunder Contract DE-AC05-76RL01830

Mongird, Kendall (ORCID:0000000328077088)↗

IM3 Projected US Data Center Locations

IM3 Projected US Data Center Locations This dataset contains model projections of new data center facilities in the contiguous United States (CONUS) through 2035 using the CERF – Data Centers model. Data center locations are modeled across four data center electricity demand growth scenarios (low, moderate, high, higher) and five market gravity scenarios (0%, 25%, 50%, 75%, 100%). Projected locations are intended to be regional representations of feasible siting locations in the future to assess potential grid and water stress impacts. The data center load growth scenarios correspond with the rates outlined in EPRI (2024) and include 3.71%, 5%, 10%, and 15% annual growth of electricity demand for data centers from 2023 values in 37 states across the CONUS. Market gravity scenarios correspond to the relative importance of proximity to data center markets or high population areas compared to locational cost in the siting algorithm. 0% market gravity means that siting decisions were entirely determined by the locational cost in each feasible location. 100% market gravity means that only market proximity was considered when siting. Other scenarios have weight placed on both components where total weight always equals 100%. Locational cost is dependent on facility cooling type and corresponding electricity cost, taxes, and other factors. Facility cooling type is spatially determined where high water stress and/or areas with high summer wet bulb temperatures are assumed to operate with mechanical cooling for a higher fraction of the year rather than evaporative cooling. Feasible data center siting areas are based on geospatial suitability raster data developed with open-source information. The following areas are excluded from siting: Areas within 300 m of a federal airport runway or within an airport area boundary Waterbodies Areas with slope >16% Areas susceptible to sinkholes High coastal or inland flood risk areas Local, state, and federal parks, leisure areas, and cemeteries Areas >2 km away from electric substations Areas >5 km away from a municipal water supplier service area Areas >2 km away from high-speed fiber provider service territory USGS Protected Areas Database of the United States (PAD-US) GAP status 1, 2, or 3 areas US National Parks Wetlands USFWS critical habitats BIA land areas Railroads, major roadways, and minor roadways Military areas and training grounds NLCD developed lands Areas >0.8 km (0.5 miles) from NLCD developed lands Because we use open-source information, proprietary information that can influence siting decisions such as individual tax agreements with cities, detailed fiber line connectivity, electric grid power capacity agreements, and others, are not currently accounted for in the modeling process. Using specific building locations and footprints in the dataset for local planning purposes is not advised. Technical Information Geospatial data is provided in geojson format using the Albers Equal Area Conic (ESRI:102003) coordinate reference system. The datasets contain the following parameters: id - unique identification number within given scenario file growth_scenario – data center demand growth scenario market_gravity_weight – market gravity weight scenario (%) region – name of region (i.e., US State) total_cost_million_usd – locational siting cost ($million) campus_size_square_ft – total land acquired for data center facility (square ft) data_center_it_power_mw – IT power of data center facility (MW) mechanical_cooling_frac – fraction of year when data center uses mechanical cooling system water_cooling_frac– fraction of year when data center uses evaporative cooling system cooling_energy_demand_mwh – total annual facility energy demand for cooling (MWh) cooling_water_demand_mgy – total annual facility water demand for cooling (MG) cooling_water_consumption_mgy – total annual facility water consumed (MG) normalized_locational_cost – normalized total locational cost score for location normalized_gravity_score – normalized market gravity score for location weighted_siting_score – total weighted siting score of locational cost and gravity score geometry – polygon geometry of facility Acknowledgment IM3 is a multi-institutional effort led by Pacific Northwest National Laboratory and supported by the U.S. Department of Energy's Office of Science as part of research in MultiSector Dynamics, Earth and Environmental Systems Modeling Program. License This data is made available under a CCBY4.0 License Disclaimer This material was prepared as an account of work sponsored by an agency of the United States Government. Neither the United States Government nor the United States Department of Energy, nor the Contractor, nor any or their employees, nor any jurisdiction or organization that has cooperated in the development of these materials, makes any warranty, express or implied, or assumes any legal liability or responsibility for the accuracy, completeness, or usefulness or any information, apparatus, product, software, or process disclosed, or represents that its use would not infringe privately owned rights. Reference herein to any specific commercial product, process, or service by trade name, trademark, manufacturer, or otherwise does not necessarily constitute or imply its endorsement, recommendation, or favoring by the United States Government or any agency thereof, or Battelle Memorial Institute. The views and opinions of authors expressed herein do not necessarily state or reflect those of the United States Government or any agency thereof. PACIFIC NORTHWEST NATIONAL LABORATORYoperated byBATTELLEfor theUNITED STATES DEPARTMENT OF ENERGYunder Contract DE-AC05-76RL01830

Mongird, Kendall (ORCID:0000000328077088)↗

Coincident Capture through Post-processing PTRAC [Slides]

This presentation discusses the new PTRAC capabilities and workflows. The PTRAC capability in MCNP6.3 has seen a massive overhaul since MCNP6.2. The new HDF5 file format allows for both MPI- and thread-based parallelism. MCNPTools has been updated to handle the new HDF5 PTRAC format and is now open sourced on GitHub. Built-in capabilities, such as the pulse-height tally coincident capture special treatment, can largely be replicated through separate postprocessing scripts that leverage both PTRAC and MCNPTools. This allows for greater flexibility in user-specified and controlled detector response functionality, ultimately using the MCNP code for what it is best at (i.e., particle transport).

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Changuinola peat soil characteristics and gas emission raw data October 2019

This dataset comprises radiocarbon and geochemical measurements from peat and porewater samples collected across various depths at a site in Bocas del Toro, Panama. The study focuses on carbon cycling dynamics in tropical peatlands by examining carbon isotopic signatures (¹⁴C and ¹³C) and elemental compositions of bulk peat, dissolved organic carbon (DOC), carbon dioxide (CO₂), and methane (CH₄). Key parameters include radiocarbon ages and isotopic ratios (δ¹³C) of bulk peat, concentrations of carbon (%C) and nitrogen (%N), and radiocarbon content of porewater gases and dissolved organic carbon (DOC). The data provide insights into the vertical and spatial distribution of carbon sources and possible preservation and decomposition processes within tropical peat profiles, offering critical information for understanding carbon storage and greenhouse gas emissions in these ecosystems.This dataset is comprised of one main data folder containing (1) file-level metadata; (2) data dictionary; (3) field metadata; (4) carbon isotopic signatures (¹⁴C and ¹³C); (5) concentrations of carbon (%C) and nitrogen (%N); (6) radiocarbon content of porewater carbon dioxide (CO₂), and methane (CH₄) ; (7) porewater DOC; (8) bulk peat sampling protocol; (9) porewater sampling protocol; (10) porewater gas collection methods; and (11) gas extraction methods. All files are in .csv format and can be opened with any software that supports this file types.

54 ENVIRONMENTAL SCIENCES↗

Microstructural dependence of defect formation in iron-oxide thin films

In this report passivating iron-oxide films are grown atop iron films simulating the corrosion process in a nuclear reactor environment. Two oxide films grown via physical vapor deposition at 600 °C and room temperature exhibited dense-epitactic and columnar-polycrystalline, microstructures respectively. A third oxide film grown in open air at 600 °C exhibited an eqiuaxed, porous morphology. Cubic maghemite and magnetite phases in each oxide film were identified via grazing incidence X-ray diffraction. Positron annihilation spectroscopy was used to characterize point defects and measure their depth and size distributions in each oxide layer and showed a range of average positron lifetimes from 0.23 ns in the high temperature, vapor deposited film, 0.35 ns in the room temperature-grown film, and 0.31 ns in the thermally grown oxide. These data indicate that the film morphology, which varies greatly in these films, leads to very different defect content. Finally, four-dimensional scanning transmission electron microscopy was used to measure the internal stress of each film and was correlated to the strain state presented in the X-ray diffraction spectra. The defect formation in each film is reasoned through using a thin film growth model.

36 MATERIALS SCIENCE↗

High Throughput Computational Framework of Materials Properties for Extreme Environments

This project aims to establish a framework capable of efficiently predicting the properties of structural materials for service in harsh environments over a wide range of temperatures and over long periods of time. The approach is to develop and integrate high throughput first-principles calculations in combination with machine learning (ML) methods, perform high throughput CALPHAD (calculations of phase diagrams) modeling, and carry out finite element method (FEM) simulations. Relevant to high temperature service in fossil power system, nickel-based superalloys such as Inconel 740 and Haynes 282 as well as the associated (Ni-Cr-Co)-Al-C-Fe-Mn-Mo-Nb-Si-Ti system, were investigated. The present framework was built on the concept of phase-based property data, in which properties of individual phases are modeled as a function of internal and external independent variables. This project established an open-source infrastructure with the following capabilities: (1) High throughput implementation of first-principles calculations at finite temperatures and variable compositions using both accurate phonon calculations and the efficient Debye model for thermodynamic properties, elastic constants, diffusion coefficients, vacancy formation, stacking and twin faults, and dislocation mobility; i.e., using the developed code DFTTK; (2) Machine learning capabilities to predict the above properties so that the number of first-principles calculations can be significantly reduced; e.g., using the developed code SIPFENN; (3) High throughput CALPHAD modeling of the above properties as a function of temperature and composition using our unique capability based on ESPEI and PyCalphad; (4) New capabilities to predict the stress-strain behavior of individual phases; and (5) New models for tensile strength prediction in common FEM software with the crystal plasticity finite element simulations (CPFEM).

, Ni-based superalloys↗

Identifying the speciation of salt-based actinides in the presence of contaminants [Slides]

Molten salt reactors (MSR) present advantages over light water reactors, such as higher safety and energy efficiency, convenient waste processing, and the ability to use more abundant thorium instead of uranium as the fuel source. However, due to the high temperatures associated with these reactors the reactor container corrosion product can have a high influence on the molten salt system, and they are considered as a part of the fuel salt system. It is essential to have a comprehensive understanding of the chemical reactions that are occurring in the molten salt in the presence of contaminants such as nickel, manganese, chromium chlorides and oxides. Because they can change the local structure of these salts and the local structures of these salts govern the thermophysical properties of the molten salts, which would determine the safety and operational parameters of the reactor. This study is focused on identifying f-element materials speciation resulting from reaction with corrosion/degradation products in a molten salt environment to understand the flow of MSR. In this work, f-element chlorides are mixed with alkali and alkali earth metals and corrosion products such as transition metal chlorides are introduced systematically, inside a glovebox. Then these are heated to around 700 ? and slowly cooled to room temperature. Afterwards, these are analyzed using different characterization techniques such as powder X-ray diffraction, UV-vis and Raman spectroscopy and solid-state NMR. The initial work was conducted with lanthanide chlorides as a surrogate for actinides and the acquired data strongly indicates that in the presence of corrosion products new phase formation/ change in coordination environments occurs. This work has been presented at multiple conference presentations. The proposed work would allow to extend this work to actinides (depleted uranium and thorium) and this would assist to identify the speciation of actinide chloride in the presence of corrosion products. Also, this work would open an opportunity to compare the coordination behavior of lanthanides with the actinides.

37 - INORGANIC, ORGANIC, PHYSICAL AND ANALYTICAL C↗

Procedure Parsing: A Method for Parsing Handwritten Documents into Computer-Based Procedures

The nuclear industry is heavily procedure driven, where almost everything has a step-by-step instruction that is expected to be followed in detail. Historically, these procedures were printed on paper copies. Recently, the industry transitioned towards electronic copies (i.e., PDFs on tablets). One major drive for this transition is the introduction of human error and loss of situation awareness when using paper copies. However, electronic copies of documents inherently have the same error traps as their paper cousins. Therefore, there is an increased interest in a way to utilize the information in the step-by-step guidance, but to present it in a dynamic manner that guides the user and adapts to any encountered conditions. Researchers at Idaho National Laboratory propose a flexible, automated method based on document parsing and augmented by natural language processing (NLP) techniques, to address these shortcomings and capitalize on these recent advancements in machine learning. The proposed method provides a cost-effective solution for computer-assisted procedure parsing of hand-written control room procedures, originally authored in Word or PDF formats, into instructions that can be displayed as computer-based procedures (CBP) in a modern graphical user interface. The researchers devised, implemented and demonstrated the Operating Procedure Extender for Novel Systems (OPENS) method in 2020. The key to OPENS is to map the original procedure text into a context-free grammar, tying content to equipment, locations, and other steps, actions, etc. This formal grammar is then used to isolate and define keywords and actions verbs, such as “measure” or “evaluate” and tie them to specific equipment referenced within that step or located in other steps, substeps, actions, subactions and tables throughout the procedure. OPENS generates an abstract syntax tree from the document which it uses to store a copy of this information in the open-standard, machine-readable and human-readable file formats XML and JSON. The XML is useful to preserve the relational aspects of the procedure for referencing tables and branching information so the user can be directed to the next appropriate active step based on the values entered for that step and previous steps. The JSON is useful for storing and exchanging data objects used to track responses to previous steps and state changes in simulated environments. In future iterations, these formats can also be used for storing more detailed information about input during plant operation or simulation. The techniques the researcher developed could further be improved by integration of recent advancements in machine learning. NLP methods could standardize documents, correct for grammatical error, and provide automated semantic validation. The researcher expects that self-supervised techniques applied to collections of natural language instructions could strengthen the model with broader context. All these methods together give us a practical way to automatically extract protocols from documents and user interactions, empowering researchers, procedure writers and nuclear operators while moving the industry forward.

99 GENERAL AND MISCELLANEOUS↗

Bay Area Regional Energy: Network Integrated Commercial Retrofits (BRICR) Project. Final Report

The BRICR project applied large-scale building energy modeling concepts with the aim of reducing the cost of energy efficiency targeting, design, and project development, and measurement of energy savings for energy efficiency programs implemented by local governments that serve small and medium commercial buildings (SMB). The project leveraged the services and resources of existing local government energy programs serving disadvantaged and hard-to-reach SMB customers. In contrast to programs run by utilities, local government programs generally do not have direct access to energy billing records for an entire class of customers in a geographic area, which prior research demonstrated useful for large-scale building energy model baseline development and calibration. , However, local governments are rich in public records that offer important clues about physical attributes and uses that, along with behavior, determine energy use. Relying only on public records, BRICR demonstrated development of credible baseline energy models for 3,792 office, retail, and hotel buildings. Publicly disclosed annual energy use data from a local energy benchmarking program and anonymized data from the Building Performance Database, the nation’s largest dataset about energy-related characteristics of buildings, were utilized to validate and calibrate energy models via an innovative method comparing distributions of energy intensity by fuel type for portfolios of buildings of similar size, vintage, and use. Portfolio calibration does not provide certainty that an energy model fits an individual building; the method is useful when billing data is not accessible – a common situation for researchers, energy service providers and ESCOs, local governments, and any party other than a utility. A software component was developed, the BRICR gem, which automates simulation when relevant data is added or edited by the user to a file saved in the standardized BuildingSync XML schema for energy audit data. The component was demonstrated as a simplified means to generate a mass of energy models corresponding to public records containing basic attributes such as building scale, location, use, year built, and aspect ratio in combination with building energy code prototype data corresponding to use and vintage. The component was also demonstrated as a simplified means to automate energy simulation when attributes are revised; the intention was to enable iterative improvement of the baseline model and energy savings estimates for common energy conservation measures as users revise relevant attributes based on their observations. In the context of institutional change and uncertainty for the participating local government energy programs, 13 whole building retrofits were completed. Impacts were measured by applying the CalTRACK2.0 methods to standardize measurement of normalized metered energy consumption. The GRIDMeter methods of stratified sampling and individual load shape analysis were applied to adjust for impacts of the effect of COVID-19 on retrofitted buildings in the context of all local buildings of similar size and use. Excluding impacts of the pandemic, retrofitted buildings demonstrated between 1.6% and 25.1% reduction in energy use. The project contributed use cases and feedback that helped inform evolution of the software tools and data formats that were combined for the first time in the BRICR project.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Seasonal enhancement of the viral shunt catalyzes a subsurface oxygen maximum in the Sargasso Sea

Subsurface oxygen maxima (SOMs) occur directly beneath the mixed layer of stratified water columns across oligotrophic open ocean basins and have been associated with physical transport processes and localized increases in phytoplankton net primary productivity (NPP). We explore the hypothesis that viral lysis (i.e., the ‘viral shunt’) increases nutrient recycling and enhances NPP, supporting SOM formation in stratified water columns, focusing on a recurring SOM at the Bermuda Atlantic Time Series (BATS) in the Sargasso Sea. Reanalysis of historical BATS data showed enhanced Prochlorococcus and virus-like particle abundances associated with SOMs. Instances of high rates of primary and secondary production observed with oxygen supersaturation further implicate a biological mechanism for SOM formation. Leveraging metatranscriptomes, metaviromes, and polony-based data collected during a Lagrangian cruise (October 2019), we link the viral shunt to SOMs, including evidence of elevated cyanophage abundance and infection of Prochlorococcus, and transcriptomic evidence of increased organic matter uptake (i.e., catabolic activity) by copiotrophic bacteria. Cruise data also showed Prochlorococcus nitrogen metabolism transcripts consistent with increased responsiveness to bacterial remineralization. These findings illustrate the biogeochemical impacts of enhanced viral lysis in marine systems, including the potential role of the viral shunt in facilitating SOM formation in the oligotrophic oceans.

Gilbert, Naomi E. [Univ. of Tennessee, Knoxville, ↗

Convective Parameters Derived from Radiosonde Data (SONDEPARAM) Value-Added Product Report

Radiosondes provide fundamental observations of the vertical profile of atmospheric state (pressure, temperature, humidity, and winds), with important implications for subsequent studies on environmental controls on cloud conditions. Within convective cloud environments, there is an increasing demand for additional value-added products (VAPs) to facilitate the use of U.S. Department of Energy Atmospheric Radiation Measurement (ARM) user facility radiosonde data sets. Such VAPs should provide quick and reliable estimates for several standard radiosonde parameters or quantities of interest using common assumptions, as well as open, flexible code for visualization and user interaction. The Convective Parameters Derived from Radiosonde Data (SONDEPARAM) VAP will apply several robust algorithms used in Wang et al. (2020) for the calculation of useful radiosonde convective cloud parameters, including the convective available potential energy (CAPE), convective inhibition (CIN), and other convective parameters, for several different assumptions regarding the initial parcel characteristics (i.e., surface-based, most unstable, mixed layer). These ARM VAP codes are developed in open, flexible Python formats, with the intention that these parameters/calculations will be incorporated into traditional ARM quick-look radiosonde plotting, yet associated with user-available codes for ease in user reproduction and assumption modification.

54 ENVIRONMENTAL SCIENCES↗

Combustion of petroleum-based transportation fuels and their blends with biofuels: a new approach for developing surrogates and understanding the effects of blending

Accurate and versatile reaction mechanisms are necessary to simulate combustion of transportation fuels blended with biofuels for improving fuel efficiency, reducing fuel consumption and mitigating the formation and emissions of combustion products harmful to the atmosphere. This project involved a combined effort of simulation and experimentation of liquid fuel burning using a combustion configuration amenable to numerical modeling with a level of detail not previously achieved for blends of biofuels with gasoline certification and surrogate fuels. The configuration of an isolated droplet burning under conditions where gas transport arises solely from fuel evaporation was selected as the platform for experiments and numerical modeling. The spherical symmetry and one-dimensional gas transport that results enabled simulating the droplet burning process with an precedented level of detail. Processes associated with unsteady gas and liquid transport, formation of particulate and gaseous products, radiation, multicomponent phase equilibrium at the droplet surface and moving boundary effects from droplet evaporation, were incorporated in a single numerical framework. The numerical model was based on the open source code OPENSmoke++ (OS) adapted to incorporate the effects noted above. In addition to simulation, an experimental design was developed for the isolated droplet configuration to acquire the data needed for validating the numerical model. Several broad accomplishments of the project were the following: demonstrating generally excellent agreement between measured and simulated combustion parameters for the fuel blends examined that included mixtures of heptane and isobutanol as a model system, and surrogates comprised of up to seven miscible components mixed with ethanol or isobutanol as representative biofuel additives; using a new approach to validate reaction mechanisms of biofuel blends which incorporated fuel evaporation into the process along with developing an experimental design for acquiring data to compare with simulations; and showing that a wealth of information could be obtained on the combustion physics of biofuels from experiments requiring volumes on the order of only nanoliters at a time thus opening the way to evaluating biofuels synthesized by new processes early in development.

02 PETROLEUM↗

Vulcan-Forge: Architecture and Design of a Multi-Modal Forensic Analysis Plugin for CALDERA

Forge and VULCAN together describe an open-architecture cybersecurity analysis ecosystem that unifies forensic artifact processing, detection engineering, and vulnerability intelligence within integrated platforms. Forge operates as a plugin for MITRE CALDERA, ingesting diverse evidence formats—including EVTX, PCAP/PCAPNG, CSV, JSON, YAML, XML, binaries, and archives—to construct a unified artifact graph enriched with severity scoring, TLP classification, and audit trails. It provides subsystems for artifact parsing, streaming structured-data visualization, NetworkMiner-based packet inspection, PE/.NET binary analysis, and LLM-assisted triage and rule generation, with outputs validated against CCCS-YARA and pySigma schemas. VULCAN complements this by serving as a cybersecurity analyst platform that integrates a Neo4j knowledge graph, Qdrant vector retrieval, SSVC-based triage, and a local LLM to deliver CVE intelligence and forensic analysis through a multi-source ingest pipeline drawing from NVD, CISA KEV, EPSS, MITRE ATT&CK, and CAPEC. Together, they bridge structured threat intelligence with automated forensic analysis and detection workflows.

97 MATHEMATICS AND COMPUTING↗