Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “software process improvement”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

Carbon dioxide, water vapor and methane soil efflux (soil respiration) in a Pinus palustris restoration site in Georgetown, SC

This dataset contains processed data from a combination of survey flux chambers and long-term automated flux chambers. Biweekly soil flux measurements were conducted from June 2023 through December 2025 at a longleaf pine restoration site in Georgetown, SC. Processed, QAQC’d data can be found in the file: 1_DATA_ESS_DOE_HR_RS_HB3_QAQC_Survey_Data_20260223.csv. Raw and working data files (.json, & .81x format) from LI-COR equipment are included for reference and can be accessed using SoilFluxPro software. CSV metadata files describe the raw data and modifications made using SoilFluxPro v5 and Matlab R2024b, as well as formatting and units for processed CSVs. Matlab code is included for reading in the processed CSVs. This research was performed as part of the project: “Improving models of stand and watershed carbon and water fluxes with more accurate representations of soil-plant-water dynamics in southern pine ecosystems”, which examines in part the effects hydraulic redistribution on soil efflux of carbon dioxide, water vapor and methane, as well as soil moisture and temperature in a southern pine ecosystem with sandy soils and high water table.

CARBON DIOXIDE FLUX↗

Systematic Benchmarking of Climate Models: Methodologies, Applications, and New Directions

As climate models become increasingly complex, there is a growing need to comprehensively and systematically assess model performance with respect to observations. Given the increasing number and diversity of climate model simulations in use, the community has moved beyond simple model intercomparison and toward developing methods capable of benchmarking a large number of simulations against a suite of climate metrics. Here, we present a detailed review of evaluation and benchmarking methods and approaches developed in the last decade, focusing primarily on scientific implications for Coupled Model Intercomparison Project (CMIP) simulations and CMIP6 results that contributed to the Intergovernmental Panel on Climate Change (IPCC) Sixth Assessment Report (AR6). Based on this review, we explain the resulting contemporary philosophy of model benchmarking, and provide clear distinctions and definitions of the terms model verification, process validation, evaluation, and benchmarking. While significant progress has been made in model development based on systematic evaluation and benchmarking efforts, some climate system biases still remain. The development of open‐source community software packages has played a fundamental role in identifying areas of significant model improvement and bias reduction. We review the key features of several software packages that have been commonly used over the past decade to evaluate and benchmark global and regional climate models. Additionally, we discuss best practices for the selection of evaluation and benchmarking metrics and for interpreting the obtained results, the importance of selecting suitable sources of reference data and accurate uncertainty quantification.

Environmental sciences↗

Evapotranspiration partitioning estimates from 8 methods from 47 NEON sites, 2019-2021

This dataset provides daily estimates of evapotranspiration (ET) and the transpiration-to-evapotranspiration ratio (T/ET) across 47 terrestrial National Ecological Observatory Network (NEON) sites spanning diverse environmental and biome conditions in the United States across three years of data (2019-2021). Daily ET is reported in both energy units (MJ m⁻² day⁻¹) and equivalent water depth (mm day⁻¹), assuming a constant latent heat of vaporization of 2.45 MJ/kg. The primary method uses a hybrid recurrent neural network–Penman–Monteith framework (RNN-PM), which integrates physically based surface energy balance constraints with data-driven learning to partition ET into transpiration and evaporation components. Model inputs include in situ meteorological observations (air temperature, vapor pressure deficit, wind speed, and radiation) combined with satellite-derived land surface temperature, leaf area index, and soil moisture. For benchmarking and uncertainty assessment, T/ET estimates from seven additional models are included: Priestley-Taylor Jet Propulsion Laboratory (PT-JPL), Penman-Monteith (P-M), Two-Source Energy Balance (TSEB), Support Vector Regression (SVR), and Categorical Boosting (CatBoost), among others—spanning empirical, machine-learning, and process-based approaches (see methods section or linked publication for detailed descriptions). Data Package Contents: The dataset a csv files containing daily ET and T/ET estimates for each site and model, along with associated metadata files these variables. Data can be accessed using common spreadsheet software (e.g., Microsoft Excel, LibreOffice) or programming environments such as R or Python. Together, these data support cross-site comparisons of ecosystem water use, evaluation of ET partitioning methods, and development of improved land–atmosphere exchange models.

EARTH SCIENCE > ATMOSPHERE↗

Determining Levels of Detail for Simulators of Parallel and Distributed Computing Systems via Automated Calibration

There are two sources of inaccuracy when simulating parallel and distributed computing systems: (i) a simulator implemented at an insufficient level of detail; and (ii) incorrectly calibrated simulation parameter values. Increasing the simulator’s level of detail can improve accuracy, but at the cost of higher space, time, and/or software complexity. Furthermore, evaluating the intrinsic accuracy of a simulator requires that its parameters be well-calibrated. Making decisions regarding the level of detail is thus challenging. We propose a methodology for instantiating the simulation calibration process and a framework for automating this process, which makes it possible to pick appropriate levels of detail for any simulator. We demonstrate the usefulness of our approach via two case studies for two different domains.

McDonald, Jessie [University of Hawaii at Manoa, H↗

FIRM image analysis: A machine learning workflow for quantifying extracellular matrix components from electron microscopy images

The extracellular matrix (ECM) is a complex network of biomolecules that plays an integral role in the structure, processes, and signaling mechanisms of cells and tissues. Identifying and quantifying changes in these matrix components provides insight into the mechanisms behind specific tissue remodeling processes; however, quantifying these changes is challenging due to difficult imaging conditions, complexity of the ECM, and the subtlety of these changes. Current imaging techniques allow us to visualize these critical remodeling events and developments in image analysis have employed a combination of analysis software and machine learning techniques to improve the efficiency and accuracy with which features are measured. Although image analysis has seen much improvement in recent years, there has been no technique developed to address ambiguity in feature edges in electron microscopy images. Presented here is a new machine learning-based workflow for the analysis of microscopy images named FIRM (Feature Identification from Raw Microscopy) that uses a random forest classifier to identify ECM features of interest and generate binary segmentation masks for quantification with ImageJ-FIJI. FIRM performed with an F1 score of 0.794 and greater than 80% accuracy for number and size of features detected. FIRM had similar deviation from the ground truth in the number of identified fibrils, fibril size, and size distributions when compared to human analyses. The results suggest that FIRM performs as well as manual analysis and requires a fraction of the time. This analysis technique is more efficient, eliminates user bias, and can be easily optimized to identify a variety of features, making it useful for any discipline requiring image analysis.

Science & Technology - Other Topics↗

Litter Production and Foliar Nutrient Resorption in Pioneer and Non-Pioneer Species in a Selective Logging Experiment in the Central Amazon, BIONTE, ZF-2, Manaus, 2022-23

This dataset was collected near the city of Manaus, Brazil, at the Experimental Station of Tropical Forestry (EEST, aka “ZF2”), inside the BIONTE (BIOmass and NuTrient Experiment). The experiment included three levels of increasing selective logging intensity, along with control, with 1-hectare permanent plots (12 total) located at the center of 4-hectare treatment plots. The vegetation has a high floristic diversity, the soils of the region are poor in nutrients, and the topography is characterized by plateaus (where BIONTE is located), and also valley bottoms and slopes. Three treatments of differing logging intensities were applied in the BIONTE experiment (T1, T2 and T3). The study was conducted in Treatment 3 (Block I – permanent plot), which represents the most intensive logging treatment, with 69% of the basal area (m²∙ha⁻¹) removed in 1988. The present dataset spans the period from May 1, 2022, to May 1, 2023. The data package includes leaf_nutrient_data, litterfall_total_data, leaf_litterfall_species_specific_data, and species_info, all provided in .csv format. These formats allow users to process and analyze the data in various software applications and programming languages, such as Python and R. This dataset was collected to advance knowledge on nutrient cycling in Amazonian forests, specifically distinguishing between species with two distinct functional traits: fast-growing and slow-growing. It also aims to improve Earth System Models, such as the E3SM Functionally Assembled Terrestrial Ecosystem Simulator (FATES). Additionally, it was used in a paper currently in preparation (Carvalho et al., in prep.), which aims to quantify seasonal litter production and foliar nutrient resorption in pioneer (fast-growing) and non-pioneer (slow-growing) tree species in the central Amazon. Specifically, it seeks to answer two key questions: 1) Is there a difference in leaf litter production, leaf nutrient flux and leaf nutrient concentration between pioneers and non-pioneers species? Is there a difference in the efficiency of foliar nutrient resorption between pioneers and non-pioneers species?

54 ENVIRONMENTAL SCIENCES↗

An Open-Source Python Package for CFD Solution Verification

Informed decision-making using computational fluid dynamics (CFD) results requires quantifying the errors and uncertainties of a simulation. Verification, validation, and uncertainty quantification (VVUQ) methods were developed to address this need and have matured. However, these VVUQ analyses are often non-trivial and require CFD analysts and practitioners to have specific skill sets. This has led to the uneven adoption of VVUQ analyses, in part, based on the availability of software tools to aid CFD analysts and practitioners. Solution verification, a procedure to evaluate the accuracy of a simulation by estimating potential errors arising from the computational model and computing the uncertainties without comparing to results from a physical system, is one of the lagging VVUQ analyses as the absence of software has forced CFD analysts and practitioners to develop their own codes or piece together incomplete software from across the internet. This work presents an opensource Python package, CFDverify, to lower the barrier of entry and fill in the technological gap in solution verification. CFDverify also provides a streamlined framework to remove some potential errors in post-processing CFD results. The hope is that CFDverify can improve the quality and quantity of CFD solution verification in scientific and research studies and attract interest in developing a communal tool. This paper describes the design, features, and an example use of CFDverify.

Weinmeister, Justin [ORNL] (ORCID:0000000160090237↗

Carbon dioxide, water vapor and methane soil efflux (soil respiration) in a Pinus palustris root exclusion in Georgetown, SC

This dataset contains processed data from a combination of survey flux chambers and long-term automated flux chambers. Soil flux measurements were conducted from June 2023 through December 2025 in a mature longleaf pine forest in Georgetown, SC. Soil respiration measurements were conducted approximately biweekly for two and a half years, before and after a root exclusion that took place on May 5, 2024. Processed, QAQC’d data for the treatment (root exclusion) and control (roots intact) before and after the root exclusion can be found in the file: 1_DATA_ESS_DOE_HR_RS_HB2_QAQC_Survey_Data_20260223.csv. Two multiday deployments were also conducted prior to the root exclusion using long-term automated chambers to continuously monitor greenhouse gas soil efflux. Processed, QAQC’d data for both long-term deployments can be found in the file: 2_DATA_ESS_DOE_HR_RS_HB2_QAQC_Longterm_Data_20260209.csv. Raw and working data files (.json, .81x, & .82z format) from LI-COR equipment are included for reference and can be accessed using SoilFluxPro software. CSV metadata files describe the raw data and modifications made using SoilFluxPro v5 and Matlab R2024b, as well as formatting and units for processed CSVs. Matlab code is included for reading in the processed CSVs, with sample figures comparing treatment and control. This research was performed as part of the project: “Improving models of stand and watershed carbon and water fluxes with more accurate representations of soil-plant-water dynamics in southern pine ecosystems”, which examines in part the effects hydraulic redistribution on soil efflux of carbon dioxide, water vapor and methane, as well as soil moisture and temperature in a southern pine ecosystem with sandy soils and high water table.

CARBON DIOXIDE FLUX↗

SolarAPP+ Enhancements and Commercialization (Final Technical Report)

With millions of distributed photovoltaic (PV) systems expected to be installed over the next five years, the permitting and inspection processes of authorities having jurisdiction (AHJs) may become overburdened, causing delays and increased costs for installed systems (Cruce et al. 2022). The central goal of this project was to automate and streamline permitting processes for distributed PV systems and complementary technologies, such as battery storage. Deploying automated permitting has been hypothesized to reduce permit review times, resulting in reduced costs and improved customer experience. This has the potential to expand the PV and PV-plus-storage market nationwide. The National Renewable Energy Laboratory (NREL) and its project partners, including UL Solutions, the Interstate Renewable Energy Council (IREC), the International Code Council (ICC), and more, developed the Solar Automated Permit Processing (SolarAPP+™) software platform to reduce permit review times. NREL and its partners also collaborated with the solar industry, the building safety community, and local governments to develop SolarAPP+.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Moving toward automated µFTIR spectra matching for microplastic identification: addressing false identifications and improving accuracy

Abstract Infrared spectroscopy is a widely used tool for studying microplastics and identifying microparticles. Researchers rely on spectral libraries to differentiate between synthetic and natural materials. Unfortunately, spectral library matching is not perfect, and best practices require researchers to use time consuming, manual peak matching to assess spectral matches. Moving toward automated matching requires increased confidence in the matching process. Using spectra matching software may increase the efficiency of particle identification, however some matching strategies may confuse natural materials such as cotton, silk, and plant matter with common classes of synthetics such as polyesters and polyamides. In this experiment, we prepared 22 pristine sample materials from natural and synthetic sources and measured micro-Fourier transform infrared (µFTIR) spectra in transmission mode for each sample using a Thermo Nicolet iN10 MX instrument. The collected spectra were then input into two spectral library matching systems (Omnic Picta and Open Specy), using a total of five identification routines. Next, we placed a subset of four pristine microplastic materials in a biologically active river system for two weeks to simulate environmental samples. These simulated environmental samples were processed using 10% hydrogen peroxide for 24 h to remove organic contamination and then identified using the strongest performing library. We found that libraries with fewer sample spectra produced lower correlation matches and that using derivative correction greatly reduced the number of inaccuracies in identifying materials as either natural or synthetic. We also found that environmental fouling reduced the correlation value of library matches when compared to pristine particles, however the effect was not consistent across the four materials tested. Overall, we found that the accuracy of automated library matching in the tested systems and processing routines varied from 64.1 to 98.0% for distinguishing between natural and synthetic materials, and that a high Hit Quality Index (HQI) did not always correlate with accuracy. These results are important for the microplastic field, demonstrating a need to rigorously test spectral libraries and processing routines with known materials to ensure identification accuracy.

Kozloski, Rachel↗

Real-time single-element-detection structured illumination optical metrology for laser powder bed fusion

Laser powder bed fusion (LPBF) is a type of metal additive manufacturing which could benefit from improved process monitoring to improve quality control. We demonstrate, for the first time to our knowledge, the coaxial monitoring of melt track formation in steel powder with spatial frequency modulation imaging (SPIFI), an enhanced-resolution imaging technique which uses a photodiode to record one-dimensional images. Using a custom live-display software and a high-speed SPIFI geometry, we offset the SPIFI field of view from the fusing beam to monitor different regions of the LPBF melt pool and surrounding area. This demonstrates the potential of SPIFI to monitor spatial features within the melt pool in real-time with increased data efficiency.

Hunter, Scott (ORCID:0009000886150312)↗

The Artificial Scientist: in-Transit Machine Learning of Plasma Simulations

Large-scale simulations or scientific experiments produce petabytes of data per run. This poses massive challenges for I/O and storage when scientific analysis workflows are run manually offline. Unsupervised deep learning-based techniques to extract patterns and non-linear relations from these large amounts of data provide a way to build scientific understanding from raw data, reducing the need for manual pre-selection of analysis steps, but require exascale compute and memory to process the full dataset available. In this paper, we demonstrate a heterogeneous streaming workflow in which plasma simulation data is streamed directly to a Machine Learning (ML) application training a model on the simulation data in-transit, completely circumventing the capacity-constrained filesystem bottleneck. This workflow employs openPMD to provide a high level interface to describe scientific data and also uses ADIOS2, to transfer volumes of data that exceed the capabilities of the filesystem. We employ experience replay to avoid catastrophic forgetting in learning from this non-steady state process in a continual manner and adapt it to improve model convergence while learning in-transit. As a proof-of-concept, we approach the ill-posed inverse problem of predicting particle dynamics from radiation in a particle-incell (PIConGPU) simulation of the Kelvin-Helmholtz instability (KHI). We detail hardware-software co-design challenges as we scale PIConGPU to full Frontier, the Top-1 system as of June 2024 Top500 list.

Kelling, Jeffrey [Helmholtz-Zentrum Dresden Rossen↗

Characterization of Infrasonic Signatures of Earth-Grazing Fireballs as Analogues to Hypersonic Vehicles (Final Report)

Accurate detection, discrimination, and characterization of high-altitude hypersonic events using infrasonic monitoring are critical to planetary defense and global strategic surveillance. This report synthesizes recent advances achieved through rigorous analysis of infrasonic signatures from natural meteoroids, emphasizing shallow entry-angle meteoroids as essentially proxies for artificial hypersonic systems. Meteoroids naturally encompass diverse velocities, trajectories, altitudes, and fragmentation behaviors, enabling systematic validation of empirical period–yield relationships, waveform morphology classifiers, and trajectory-induced back-azimuth deviation models. Integration of adaptive array-processing enhancements within Cardinal software further extends infrasonic detection sensitivity and signal classification reliability. Collectively these advances, based solely on infrasonic signatures or limited optical data, offer robust methodologies for distinguishing natural from artificial hypersonic sources, significantly reducing event geolocation uncertainties and refining source-function determination. The outcomes detailed herein lay foundational groundwork for improved global hypersonic event-surveillance frameworks, supporting improved security preparedness and informing strategic monitoring and defense policies.

54 ENVIRONMENTAL SCIENCES↗

Vision and Development of a Design, Implementation, and Verification Automation (DIVA) Software Platform for DNA Construction

Abstract DNA construction, while a prerequisite to many biological endeavors, is often a time-consuming distraction from an individual’s primary research objectives. We envisioned that with the right software infrastructure and cultural mindset, a single person could execute in parallel the batched DNA construction tasks of an entire research institute, at scales realizing efficiency gains through process and laboratory automation. In pursuit of this vision, we developed the Design, Implementation, and Verification Automation (DIVA) software platform. DIVA’s web interface enables researchers to design DNA constructs (using visual biological computer-aided design tools and biological parts repositories), submit designs for construction to dedicated staff, and track DNA construction as it progresses. DIVA supports the dedicated staff through the DNA construction process and records both successful and unsuccessful attempts toward improving the overall process. The platform is publicly available at public-diva.jbei.org and its open-source code through github.com/JBEI/DIVA.

Plahar, Hector [DOE Agile BioFoundry , , ,; DOE Jo↗

PyTUQ: Python Toolkit for Uncertainty Quantification

SAND2025-03661O PyTUQ is a user-friendly software toolkit designed to help researchers and professionals understand and manage uncertainty in several scientific fields. By providing tools for analyzing how uncertainties affect outcomes, PyTUQ can be applied in areas such as energy production, and biology. Its unique approach allows users to make more informed decisions by assessing risks and improving predictions. Whether you're studying combustion processes or exploring complex biological systems, PyTUQ empowers you to gain deeper insights and enhance the reliability of your results. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

SciDAC↗

Novel technology of non-contact real-time radiation damage sensors for high power targets.

This report summarizes the contributions of an intern participating in the Community College Internship (CCI) program at Fermilab, focusing on the development of a novel, non-contact, real-time radiation damage sensor technology. The core objective is to create a reliable sensor capable of measuring radiation-induced degradation on high-power targets without physical contact. The experiment involves using a Class 3B supercontinuum laser directed toward a single material sample placed within a vacuum test chamber. The laser beam reflects off the sample's surface, with changes in reflectivity, indicative of radiation damage, measured by a spectrometer positioned at the chamber’s output port. The intern’s primary responsibilities included designing an interlock system to ensure laser operational safety, developing a camera-based monitoring system using Raspberry Pi devices, and creating structural supports using 3D modeling and printing techniques. Components for the interlock and camera systems were successfully designed and ordered, with preliminary 3D models printed and refined through iterative testing. Challenges encountered in the 3D printing process, such as fragile initial prototypes and difficult support removal, were overcome by adjusting printer settings and incorporating design enhancements like chamfered edges. Future activities, pending component delivery, involve installing and configuring the interlock and camera systems, as well as further improving the structural supports. Overall, the internship significantly enhanced the intern’s technical proficiency in hardware design, software integration, and advanced 3D printing, contributing directly to Fermilab’s operational safety standards and experimental effectiveness in high-energy physics research.

Pumarino Meza, Rafael [Unlisted; Fermilab]↗

Capturing Secondary Kinetic Instabilities in Three‐Dimensional Dayside Reconnection Using an Improved Gradient‐Based Closure

Magnetic reconnection is a highly dynamic process that excites a wide variety of kinetic waves and instabilities. Transverse current sheet instabilities such as the lower-hybrid drift and secondary drift-kink instabilities in particular have been shown by kinetic simulations to modify the reconnection and introduce significant turbulence and mixing to the reconnection layer. Past studies using the ten-moment fluid model to capture important kinetic physics such as the electron inertia and full representation of the pressure tensor proved advantageous to a two-fluid representation of reconnection, but the model struggled when using a local relaxation closure for the heat flux to replicate the current sheet instabilities and subsequent mixing seen in kinetic simulations. This work uses the Gkeyll software framework to perform simulations of asymmetric reconnection based on the 16 October 2015 MMS crossing of a diffusion region, the Burch event. An improved gradient-based heat flux closure is implemented, showing significant improvement in secondary kinetic instabilities that grow in the current sheet. These instabilities generate turbulence which leads to growth of secondary magnetic islands and flux ropes.

Bradshaw, K. [Princeton University, NJ (United Sta↗

Generic and ML Workloads in an HPC Datacenter: Node Energy, Job Failures, and Node-Job Analysis

HPC datacenters offer a backbone to the modern digital society. Increasingly, they run Machine Learning (ML) jobs next to generic, compute-intensive workloads, supporting science, business, and other decision-making processes. However, understanding how ML jobs impact the operation of HPC datacenters, relative to generic jobs, remains desirable but understudied. In this work, we leverage long-term operational data, collected from a national-scale production HPC datacenter, and statistically compare how ML and generic jobs can impact the performance, failures, resource utilization, and energy consumption of HPC datacenters. Our study provides key insights, e.g., ML-related power usage causes GPU nodes to run into temperature limitations, median/mean runtime and failure rates are higher for ML jobs than for generic jobs, both ML and generic jobs exhibit highly variable arrival processes and resource demands, significant amounts of energy are spent on unsuccessfully terminating jobs, and concurrent jobs tend to terminate in the same state. We open-source our cleaned-up data traces on Zenodo (https://doi. org/10.5281/zenodo.13685426), and provide our analysis toolkit as software hosted on GitHub (https://github.com/atlarge-research/2024-icpads-hpc-workload-characterization). This study offers multiple benefits for data center administrators, who can improve operational efficiency, and for researchers, who can further improve system designs, scheduling techniques, etc.

crossanalysis↗