Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Compiler techniques”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Comprehensive earthquake catalogue update and spatiotemporal distribution analysis for Iraq and surrounding regions, northeastern Arabian Plate

The updated earthquake catalogue for Iraq covers the period from 1900 to the end of 2021 and includes over 37 000 recorded earthquakes. To create this comprehensive catalogue, five key steps were taken: compiling bulletins, calculating moment magnitudes, harmonizing magnitudes, establishing empirical conversion relations and evaluating the completeness of the catalogue. A notable enhancement in this update is the direct calculation of moment magnitudes for approximately 2800 earthquakes, achieved through the coda envelope technique and waveform data from the Mesopotamian Seismological Network (MPSN) in Iraq. This updated catalogue serves as a valuable resource for examining the spatiotemporal distribution of earthquakes, with respect to earthquake density, maximum moment magnitude and seismogenic depths. Additionally, the Gutenberg–Richter relationship was applied to calculate the a- and b-values specific to Iraq. The findings show that the Zagros Fold-Thrust Belt has a seismogenic layer (source) that ranges from 2 to 33 km deep and experiences high seismic activity. In contrast, the Mesopotamian Foredeep has a seismogenic layer ranging from 1 to 25 km deep and has lower seismic activity. The greatest seismic activity is concentrated around the Mandili-Badra-Teeb fault, which has experienced significant ruptures over time. The Outer Arabian Platform is identified as the main area of seismic activity, while additional activity occurs on the Inner Arabian Platform. Three major tectonic boundaries define the distribution of earthquakes in the northeastern Arabian Plate. These boundaries are defined by the Main Zagros Reverse Fault, the Zagros Foredeep Fault and the Anah Graben and Abu Jir-Euphrates Fault Zone. These boundaries highlight variations in seismicity levels and the spatial distribution of deformation in the region. The updated earthquake catalogue presented in this study is expected to play a vital role in regional seismicity assessments and seismic hazard analyses for Iraq and its surrounding areas.

58 GEOSCIENCES↗

Using modularity to segment binary code

We consider the problem of recovering program structure from compiled binary code. We first extract the call graph and layout of functions in memory from the compiled code and represent this information in a graphical format. We then employ Louvain's modularity algorithm to identify clusters of functions that are considered to be related. We find that the quality and properties of clusters extracted by our technique are greatly impacted by the relative importance we assign to the call graph and the ordering of functions in memory.

97 MATHEMATICS AND COMPUTING↗

The Applicability of Unit Systems to High-Performance Computing Applications

Dimensional analysis is a key technique used to verify the soundness of scientific models. Most experts agree that engineering and scientific software would be made more reliable by integrating dimensional analysis in their type system. We explored how High Performance Computing (HPC) applications could integrate compile-time dimensional analysis. We started by investigating various implementation of unit systems for C++. Eventually, selecting the latest (and most advanced) one to apply to our test codes. We worked with code of increasing complexity, from a projectile trajectory calculation to the proxy-application Lulesh. This included our code, Springs-3D, which focuses on demonstrating language features while performing simple physic computations. Finally, our main contribution is a source-code analysis which extracts constraints on the dimension of all variables, functions, and constants in an application. This resulting system of equations is solved using the dimensions of a few of these objects. This analysis has the potential to greatly reduce the time spent performing dimensional analysis when refactoring application to use a representation of units.

97 MATHEMATICS AND COMPUTING↗

Optical spectroscopy of molten fluorides: Methods, electronic and vibrational data, structural interpretation, and relevance to radiative heat transfer

In this study, to help address the need for predicting radiative heat transfer (RHT) behavior of molten salts, we conducted a comprehensive review of methods and data from optical spectroscopic measurements on molten fluoride salts. Transmittance, reflectance, and trans-reflectance experimental methods are discussed, along with the corresponding data reduction methodology and the limitations of each technique. Optical spectroscopy is a convenient indirect probe for changes in structural parameters with temperature and composition. Electronic and vibrational absorption data for transition-metal, lanthanide, and actinide solutes and vibrational absorption data for alkali and alkaline earth fluoride solvents are compiled, and the corresponding structural interpretation is discussed and compared with other experimental and theoretical work. We find that solvent and solute vibrational absorption can be significant in the mid-infrared, resulting in near-infrared edges of significance to RHT. Extrapolation and averaging of existing edge data leads to estimated gray absorption coefficient values at 700 °C of 546 m —1 for FLiBe and 276 m —1 for FLiNaK, both within the range of 1 – 6000 m —1 identified to be of engineering relevance for radiative heat transfer analysis.

36 MATERIALS SCIENCE↗

Signatures of muonic activation in the Majorana Demonstrator

Experiments searching for very rare processes such as neutrinoless double-beta decay require a detailed understanding of all sources of background. Signals from radioactive impurities present in construction and detector materials can be suppressed using a number of well-understood techniques. Background from in situ cosmogenic interactions can be reduced by siting an experiment deep underground. However, the next generation of such experiments have unprecedented sensitivity goals of 10 28 years half-life with background rates of 10 -5 cts/(keV kg yr) in the region of interest. To achieve these goals, the remaining cosmogenic background must be well understood. In the work presented here, Majorana Demonstrator data are used to search for decay signatures of metastable germanium isotopes. Contributions to the region of interest in energy and time are estimated using simulations and compared to Demonstrator data. Correlated time-delayed signals are used to identify decay signatures of isotopes produced in the germanium detectors. A good agreement between expected and measured rate is found and different simulation frameworks are used to estimate the uncertainties of the predictions. The simulation campaign is then extended to characterize the background for the LEGEND experiment, a proposed tonne-scale effort searching for neutrinoless double-beta decay in 76 Ge .

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

PANTHER: A Programmable Architecture for Neural Network Training Harnessing Energy-Efficient ReRAM

The wide adoption of deep neural networks has been accompanied by ever-increasing energy and performance demands due to the expensive nature of training them. Additionally, numerous special-purpose architectures have been proposed to accelerate training: both digital and hybrid digital-analog using resistive RAM (ReRAM) crossbars. ReRAM-based accelerators have demonstrated the effectiveness of ReRAM crossbars at performing matrix-vector multiplication operations that are prevalent in training. However, they still suffer from inefficiency due to the use of serial reads and writes for performing the weight gradient and update step. A few works have demonstrated the possibility of performing outer products in crossbars, which can be used to realize the weight gradient and update step without the use of serial reads and writes. However, these works have been limited to low precision operations which are not sufficient for typical training workloads. Moreover, they have been confined to a limited set of training algorithms for fully-connected layers only. To address these limitations, we propose a bit-slicing technique for enhancing the precision of ReRAM-based outer products, which is substantially different from bit-slicing for matrix-vector multiplication only. We incorporate this technique into a crossbar architecture with three variants catered to different training algorithms. To evaluate our design on different types of layers in neural networks (fully-connected, convolutional, etc.) and training algorithms, we develop PANTHER, an ISA-programmable training accelerator with compiler support. Our design can also be integrated into other accelerators in the literature to enhance their efficiency. Our evaluation shows that PANTHER achieves up to 8.02×, 54.21×, and 103× energy reductions as well as 7.16×, 4.02×, and 16× execution time reductions compared to digital accelerators, ReRAM-based accelerators, and GPUs, respectively.

42 ENGINEERING↗

Agile Acceleration of LLVM Flang Support for Fortran 2018 Parallel Programming

The LLVM Flang compiler ("Flang") is currently Fortran 95 compliant, and the frontend can parse Fortran 2018. However, Flang does not have a comprehensive 2018 test suite and does not fully implement the static semantics of the 2018 standard. We are investigating whether agile software development techniques, such as pair programming and test-driven development (TDD), can help Flang to rapidly progress to Fortran 2018 compliance. Because of the paramount importance of parallelism in high-performance computing, we are focusing on Fortran’s parallel features, commonly denoted “Coarray Fortran.” We are developing what we believe are the first exhaustive, open-source tests for the static semantics of Fortran 2018 parallel features, and contributing them to the LLVM project. A related effort involves writing runtime tests for parallel 2018 features and supporting those tests by developing a new parallel runtime library: the CoArray Fortran Framework of Efficient Interfaces to Network Environments (Caffeine).

Rasmussen, Katherine↗

FY 2026 Midyear Report: Seismic Monitoring of Underground Vibration Sources Using Distributed Acoustic Sensing and Seismometers

Safeguards-relevant temporal changes in underground facilities can be observed using geophysical monitoring techniques. Seismic waves, in particular, provide valuable insights into subsurface activities and can serve as an important tool for detecting anomalous events that may indicate containment breaches at geological repositories. This midyear report summarizes ongoing efforts to automatically and rapidly detect and locate anomalous vibration signals that could be indicative of potential containment breaches. Previous work during FY25 focused on compiling continuous seismic datasets from two underground sites and developing a database of continuous waveforms and ground-truth event data derived from multiple sensing modalities. Building on this foundation, we are adapting anomaly detection and geolocation algorithms to explore methods for monitoring underground activities using two relatively low-maintenance sensing technologies: a dense surface geophone array deployed at the Pleasant Gap mine in Pennsylvania, and a three-dimensional fiber-optic cable array for distributed acoustic sensing (DAS) installed in the subsurface at the Sanford Underground Research Facility (SURF) in South Dakota. This report summarizes work conducted during the first two quarters of FY26, during which we refined a dynamic power spectral density (PSD)-based detector, applied it independently to each geophone station, and then combined the per‑station detections with density-based spatial clustering of applications with noise (DBSCAN) to cluster events and produce spatial maps over a nine‑day interval. In addition, we outline plans for a field trial at the Waste Isolation Pilot Plant (WIPP) in New Mexico to compare traditional seismic monitoring approaches with DAS techniques and to evaluate the benefits of combined data analysis. Activities during the past two quarters have included the preparation and submission of a Field Test Plan to WIPP for approval, as well as submission to headquarters for review and feedback.

58 GEOSCIENCES↗

X-ray diffraction under grazing incidence conditions

Material properties frequently relate to structures at or near surfaces, particularly in thin films. As a result, it is essential to understand these structures at the molecular and atomistic levels. The most accurate and widely used techniques for characterizing crystallographic order are based on X-ray diffraction. When dealing with thin films or interfaces, standard approaches, such as single crystal or powder diffraction, are not suitable. However, X-ray diffraction under grazing incidence conditions can provide the required information. Here, in this Primer, grazing incidence X-ray diffraction (GIXD) is comprehensively introduced, starting from basic considerations on X-ray diffraction at crystals with reduced dimensionality and the optical properties of X-rays, followed by a more in-depth description of an experimental performance, including X-ray sources, goniometers and detectors. Experimental errors, limitations and reproducibility are discussed. Various applications, from highly ordered inorganic single crystal surfaces to weakly ordered polymer thin films, are presented to illustrate the potential of GIXD. Data visualizations, representations and evaluation strategies are summarized, based on the example of anthracene thin films. The Primer compiles information relevant to perform high-quality GIXD experiments, evaluate data and interpret results, to extend knowledge about X-ray diffraction from surfaces, interfaces and thin films.

36 MATERIALS SCIENCE↗

Permafrost Region Greenhouse Gas Budgets Suggest a Weak CO 2 Sink and CH 4 and N 2 O Sources, But Magnitudes Differ Between Top-Down and Bottom-Up Methods

Large stocks of soil carbon (C) and nitrogen (N) in northern permafrost soils are vulnerable to remobilization under climate change. However, there are large uncertainties in present-day greenhouse gas (GHG) budgets. We compare bottom-up (data-driven upscaling and process-based models) and top-down (atmospheric inversion models) budgets of carbon dioxide (CO 2 ), methane (CH 4 ) and nitrous oxide (N 2 O) as well as lateral fluxes of C and N across the region over 2000–2020. Bottom-up approaches estimate higher land-to-atmosphere fluxes for all GHGs. Both bottom-up and top-down approaches show a sink of CO 2 in natural ecosystems (bottom-up: -29 (-709, 455), top-down: -587 (-862, -312) Tg CO 2 -C yr -1 ) and sources of CH 4 (bottom-up: 38 (22, 53), top-down: 15 (11, 18) Tg CH 4 -C y -1 ) and N 2 O (bottom-up: 0.7 (0.1, 1.3), top-down: 0.09 (-0.19, 0.37) Tg N 2 O-N yr -1 ). The combined global warming potential of all three gases (GWP-100) cannot be distinguished from neutral. Over shorter timescales (GWP-20), the region is a net GHG source because CH 4 dominates the total forcing. The net CO 2 sink in Boreal forests and wetlands is largely offset by fires and inland water CO 2 emissions as well as CH 4 emissions from wetlands and inland waters, with a smaller contribution from N 2 O emissions. Priorities for future research include the representation of inland waters in process-based models and the compilation of process-model ensembles for CH 4 and N 2 O. Discrepancies between bottom-up and top-down methods call for analyses of how prior flux ensembles impact inversion budgets, more and well-distributed in situ GHG measurements and improved resolution in upscaling techniques.

54 ENVIRONMENTAL SCIENCES↗

Advanced Laboratory and Field Arrays (ALFA)/Lab Collaboration Project (LCP) for Marine Energy (Final Scientific/Technical Report)

The objective of the Advanced Laboratory and Field Arrays (ALFA) project was to reduce the Levelized Cost of Energy (LCOE) of Marine and Hydrokinetic (MHK) energy by leveraging research, development, and testing capabilities at Oregon State University, University of Washington, and the University of Alaska, Fairbanks. ALFA is a project within the Pacific Marine Energy Center (PMEC; formerly NNMREC), a multi-institution entity with a diverse funding base that focuses on research and development for marine renewables. The ALFA project aimed to accelerate the development of next-generation arrays of wave energy conversion (WEC) and tidal energy conversion (TEC) devices through a suite of field-focused R&D activities spanning a broad range of strategic opportunity areas identified in the Funding Opportunity Announcement: • Device and/or array operation and maintenance (O&M) logistics development; • High-fidelity resource characterization and/or modeling technique development and validation; • Array-specific component technology development (e.g. moorings and foundations, transmission, and other offshore grid components); • Array performance testing and evaluation; and • Novel cost-effective environmental monitoring techniques and instrumentation testing and evaluation. The objective of the Lab Collaboration Project (LCP) was to accelerate the development of next-generation marine energy conversion systems. The LCP aimed to achieve these project objectives in collaboration with the national laboratories by: • Developing concept generation and assessment tools; • Improving access to existing testing resources; • Validating collision risk models between fish and turbines; and • Advancing analysis and simulation capabilities for wave-WEC interactions and PTO analysis in nonlinear ocean waves. The ALFA portion of the project was comprised of six overarching technical tasks: • Task 1: Debris Modeling, Detection and Mitigation; • Task 2: Autonomous Monitoring & Intervention; • Task 3: Resource Characterization for Extreme Conditions; • Task 4: Robust Models for Design of Offshore Anchoring and Mooring Systems; • Task 5: Performance Enhancement for Marine Energy Converter (MEC) Arrays; and • Task 6: Evaluating Sampling Techniques for MHK Biological Monitoring. The LCP was divided into four overarching technical tasks: • Task 7: Project Management and Reporting • Task 8: Novel Design and Assessment Methodologies for Wave Energy Converter Design (Wave- SPARC) • Task 9: Testing Access for Commercial Marine Renewable Energy Technology Developers • Task 10: Quantifying Collision Risk for Fish and Turbines • Task 11: Nonlinear Ocean Waves and PTO Control Strategy Each ALFA/LCP task listed above functioned as a separate and discreet project. A final Technical Report was written for each individual task and these reports were uploaded to OSTI, after receiving DOE approval. The following document is a compilation of each of these final, approved reports arranged as individual chapters.

13 HYDRO ENERGY↗

Automated construction of clear-sky dictionary from all-sky imager data

All-sky imagers (ASIs) have significant promise as scalable sensors for short-term solar irradiance forecasting. Many of the current computational techniques that use ASIs for this purpose rely on collections of clear-sky images indexed by time of day, solar angle, or both, called clear-sky dictionaries (CSDs). These CSDs act as baselines against which images can be compared to locate and classify clouds within the image frame. CSDs are often compiled by hand, where individuals visually inspect collections of images one at a time to find clear-sky images. This process is not scalable, and it is prone to error. This paper proposes an automated, nonparametric alternative that uses the principles of digital image processing to find clear-sky images within a set of images taken over several days. We use ground-truth measurements of the clearness index to assess the performance of our method, and we show that the images it selects accurately correspond to clear-sky images. We also compare our proposal, which is nonparametric, with a state-of-the-art parametric method. The numerical results indicate that the performance of the method proposed here is superior.

14 SOLAR ENERGY↗

Data and scripts associated with a manuscript on a meta-analysis synthesizing stream biogeochemical response to wildfires across space and time (v2)

This data package is associated with the publication “Catchment characteristics modulate the influence of wildfires on nitrate and dissolved organic carbon in lotic systems across space and time: A meta-analysis” submitted to Global Biogeochemical Cycles (Cavaiani et al. 2025). This study uses meta-analytical techniques to evaluate the effect of wildfire on in-stream responses in burned and unburned watersheds. The study aims to provide additional insight into the range of responses and net influences that wildfires have on hydro-biogeochemistry across broad spatial scales, burn extents, and the persistence of water-quality change. This study compiles data and metadata from 18 total publications that includes 1) surface water geochemistry data (dissolved organic carbon; nitrate), 2) climate classifications, 3) year of the wildfire, 4) the time lag between when the fire occurred and when the sampling occurred, and 5) study design of the publication. In total, this meta-analysis draws data that spans 8 climate guilds, 3 biomes, 62 watersheds, and 20 unique wildfires. See Sites_meta_data.csv for citations of the papers used in this meta-analysis. All R scripts and the associated data can also be found on GitHub at This data package was originally published in March 2024. It was updated in April 2025 (v2; new and modified files). See the change history section in the readme for more details. This data package contains five primary folders that include the following: (1) inputs; (2) output for analysis; (3) initial plots; (4) R scripts; and (5) GIS data. The data package also contains a data dictionary (dd) that provides column header definitions and a file-level metadata (flmd) file that describes every file. The “inputs” folder contains a list of all publications identified during the formal web search and an indication of whether each publication was included in the final analysis. Additionally, it includes site-level metadata, catchment characteristics, and GIS data for all publications included in the final analysis. The “Output_for_analysis” folder contains all data frames and figures generated from each R script used for additional data analysis. The “initial_plots” folder includes all exploratory figures that will be included in a supplemental and figures that will be submitted with the manuscript for publication. The “R_scripts” folder contains the scripts that perform all the data manipulations, statistical analyses, and plots. The “gis_data” folder includes shape files for each fire included in this meta-analysis. This data package contains the following file types: csv, pdf, jpeg, cpg, dbf, prj, shp, shp.ea.iso.xml, shp.iso.xml, shx.

54 ENVIRONMENTAL SCIENCES↗

Agentic AI vs ML-Based Autotuning: A Comparative Study for Loop Reordering Optimization

High Performance Computing (HPC) applications rely heavily on code optimizations to achieve good performance on modern CPU and GPU architectures. Traditional Machine Learning auto-tuning approaches have demonstrated success in exploring high-dimensional spaces, but they often require expensive compile-run evaluations and lack adaptability for large HPC applications. The recent advances in Large Language Models (LLMs) and Agentic AI systems raise intriguing questions about the potential of these approaches to address specific optimization methodologies. This work aims to answer an essential question for the HPC community: “How Agentic AI Systems Compare to Traditional ML Autotuning Techniques?” To address this question, we present a comparative analysis between a traditional ML-based optimization approach and an Agentic AI system, evaluating their respective capabilities and limitations for loop-level optimization. In addition, we introduced a new Agentic AI system named LoopGen-AI using three different Large Language Models: GPT-4.1, Claude 4.0, and Gemini 2.5. A key finding is that LoopGen-AI achieves competitive per-formance with only a few program runs, the reasoning logs from the agents revealed that their decisions rely heavily on the combination of semantic understanding of the target kernel with dynamic feedback from the environment, highlighting a promising new dimension in performance tuning. In contrast, ML-based autotuners focus on statistical exploration, and require orders of magnitude more runs to reach peak performance. Additionally, our analysis shows that prompt engineering, particularly using Persona + Context Manager patterns, significantly impacts the effectiveness of Agentic AI. Our results indicate that while Agentic AI systems are not yet a complete replacement for ML-based autotuners, it can effectively complement traditional methods.

Rosas, Miguel Romero↗

Automated electrosynthesis reaction mining with multimodal large language models (MLLMs)

Leveraging the chemical data available in legacy formats such as publications and patents is a significant challenge for the community. Automated reaction mining offers a promising solution to unleash this knowledge into a learnable digital form and therefore help expedite materials and reaction discovery. However, existing reaction mining toolkits are limited to single input modalities (text or images) and cannot effectively integrate heterogeneous data that is scattered across text, tables, and figures. In this work, we go beyond single input modalities and explore multimodal large language models (MLLMs) for the analysis of diverse data inputs for automated electrosynthesis reaction mining. We compiled a test dataset of 65 articles (MERMES-T24 set) and employed it to benchmark five prominent MLLMs against two critical tasks: (i) reaction diagram parsing and (ii) resolving cross-modality data interdependencies. The frontrunner MLLM achieved ≥96% accuracy in both tasks, with the strategic integration of single-shot visual prompts and image pre-processing techniques. We integrate this capability into a toolkit named MERMES (multimodal reaction mining pipeline for electrosynthesis). Our toolkit functions as an end-to-end MLLM-powered pipeline that integrates article retrieval, information extraction and multimodal analysis for streamlining and automating knowledge extraction. This work lays the groundwork for the increased utilization of MLLMs to accelerate the digitization of chemistry knowledge for data-driven research.

Leong, Shi Xuan↗

Evaluation of Hardware and Software Bill of Materials (HBOMs/SBOMs) Extraction Methods

Hardware and software bills of materials (HBOMs and SBOMs) provide important visibility into the components, dependencies, and supply chain relationships within programmable digital devices. This visibility is critical for advanced nuclear reactor applications, where use of common or shared hardware components, software libraries, suppliers, or manufacturing processes may create common cause failure (CCF) vulnerabilities despite apparent diversity. This paper evaluates current approaches for obtaining and analyzing HBOMs and SBOMs in support of CCF, diversity and defense-in-depth (D3) assessments, and begins to explore potential methods for artificial intelligence/machine learning-based analysis. The availability of BOM information from advanced reactor manufacturers and vendors, representative hardware and software categories found in advanced reactor systems continues to limit research [13]. This paper compares commonly used BOM formats, including CycloneDX, SPDX, and SWID. It also surveys publicly available tools for generating BOMs from source code, compiled binaries, and hardware-related information, noting limitations in language coverage, system age, and format interoperability. Finally, this paper evaluates methods for correlating BOM data with vulnerability and exploitability information, including VEX, CVE, and CWE resources. The findings indicate that publicly available nuclear-vendor BOMs are limited, making third-party extraction and research into novel analysis techniques necessary.

Cybersecurity↗

HVAC and Control Templates for the Modelica Buildings Library

This article reports on our experience in creating Modelica models for systems with thousands of configurations and closed-loop controls. The development of such templates required exploration of class parameterization techniques and data structures for handling large sets of equipment parameters. By describing these issues and the approach taken, we show how the Modelica language can support advanced templating logic. The main limitation we encountered relates to parameter assignment and propagation. The interpretation of parameter attributes at user interface runtime, or the handling of non-trivial constructs involving record classes at compile time is not consistently supported by Modelica tools. This leads to choices that are difficult to make when looking for a generic implementation.

Gautier, Antoine↗