Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “data cube”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

A Taxonomic Classification Approach for Global Spatio-temporal Data

The World Bank, World Health Organization, and other major vendors collectively provide thousands of global time series datasets that focus on issues of the environment, public health, economics, violence, education, and national security. Sorting these data into meaningful information requires the use of data mining techniques to cluster trends into an orderly and manageable number of cases. The World SpatioTemporal Analytics and Mapping (WSTAMP) project database (wstamp.ornl.gov) was developed to spatiotemporally harmonize global vendor data (23,300+ attributes, 200+ locations, 50+ years). Within the WSTAMP analytical environment, Dynamic Time Warping (DTW) has been a highly effective data-driven approach for clustering and mapping these time series into national spatiotemporal behavior maps. Two significant properties have surfaced from this work. First, several recognizable cluster patterns have emerged and persist across a range of locations, attributes, and time frames (e.g., increasing, decreasing, rebounding, peak, oscillating). Secondly, practitioners engaging WSTAMP have noted the explanatory and anticipatory value of these patterns and articulated particular interest in detecting them within the spatiotemporal cube. This need was addressed by shifting DTW-based clustering from an open ended, data-driven implementation to a taxonomic pattern matching approach. This paper presents the method including implementation strategies for visualization and human computer interaction and applies the approach to a sample data set and concludes with next steps.

Stewart, Robert↗

HostSub_GP: Precise Galaxy Background Subtraction in Transient Long-slit Spectroscopy with Gaussian Processes

We present a novel host galaxy subtraction technique in long-slit spectroscopy for extragalactic transients. Unlike classic methods which generally estimate the background using simple interpolation of local galaxy flux in the 2D spectrum, our approach leverages multi-band archival images of the host galaxies to model the background emission from the galaxy in the 2D spectrum. Such imaging encodes the wavelength-dependent galaxy profile along the slit, and is readily accessible through wide-field imaging surveys. We construct a smooth prior for the 2D galaxy profile with a Gaussian process (GP) based on these reference images, and use another GP to model the correlated deviations from the prior in the observed spectrum. This enables accurate inference of the galaxy flux blended with the transient. On synthetic long-slit data of a spiral galaxy extracted from a Multi Unit Spectroscopic Explorer hyper-spectral cube, the GP method remains robust as long as the host galaxy is spatially resolved and consistently outperforms classic methods. We apply the method to archival Keck spectra of two real transients, SN 2019eix and AT 2019qiz, to further demonstrate how the method uniquely recovers weak spectral features amid strong galaxy contamination, enabling refined constraints on the properties of both transients. We have released the software implementation, HostSub_GP, a scalable toolkit that leverages JAX, with an MIT license.

79 ASTRONOMY AND ASTROPHYSICS↗

Automatic Point Cloud Building Envelope Segmentation (AutoCuBES)

The Auto-CuBES algorithm is based on unsupervised machine learning that automatically labels 3D point cloud data and reduces the time spent in manual segmentation. The algorithm can process high-resolution point clouds and generate a wire-frame building envelope model with a small set of calibration parameters. The algorithm inputs a 3D point cloud generated by commonly available surveying equipment and outputs a wire-frame model of the building envelope. Unsupervised machine learning methods were used to identify facades, windows, and doors while minimizing the number of calibration parameters.

Puente, BryanMaldonado↗

Chloroform Fumigation Extraction for Microbial Biomass and Dissolved Organic Carbon from SPRUCE, Marcell Experimental Forest, Minnesota, 2021, 2022, and 2024

This data set provides the results for chloroform fumigation extraction (CFE) of peat samples collected from ambient and experimental plots in the Spruce and Peatland Responses Under Environmental Change (SPRUCE) Experiment site in June and August of 2021, June of 2022, and June, August, and October of 2024. The SPRUCE Experiment site is in the Marcell Experimental Forest in northern Minnesota, USA. The data set includes values for microbial biomass carbon (MBC), microbial biomass nitrogen (MBN), dissolved organic carbon (DOC), dissolved nitrogen (DN), moisture content (MC, available for 2021 and 2022 only) and gravimetric water content (GWC) at 11 depth increments of two-meter peat cores taken from 12 sampling sites at SPRUCE (10 temperature treatment enclosures, 2 ambient temperature treatment enclosures). The sample analysis followed standard methods. The samples were analyzed using a Shimadzu Total Organic Carbon/Nitrogen (TOC/N) analyzer (TOC-V and TOC-L; 2021-2022) or an Elementar vario TOC Cube (2024), liquid catalytic oxidation combustion analyzers for total carbon and nitrogen analysis. This dataset contains two data files in comma separate (.csv) format. Additional metadata are provided: two data dictionaries and a file-level metadata file in comma separate (.csv) format and a user guide in PDF (*.pdf) format.

dissolved nitrogen↗

High strain-rate strength response of single crystal tantalum through in-situ hole closure imaging experiments

The properties of crystalline materials often depend on directionality and operating conditions. Specifically, the strength of materials can depend anisotropically on crystal direction and the loading condition. To probe these effects, a preliminary series of high strain-rate (> 105/s) strength plate-impact hole closure experiments were performed on high purity single crystal Tantalum cubes. The orientation of the single crystals with respect to impact/loading were varied to provide data to inform crystal plasticity modeling efforts. The experiments consist of in-situ high-resolution X-ray radiographic imaging of the hole collapse under dynamic compression conditions to infer the material strength via its resistance to closure at increasing levels of plastic strain. The experiments are compared against hydrocode simulation predictions. Here, a comparison with simple elastic perfectly plastic strength model predictions is presented to elucidate the response of the different crystal orientations at high strain-rate and large plastic strains.

36 MATERIALS SCIENCE↗

Gauging nexus between topological and fracton phases

Coupled layer constructions are a valuable tool for capturing the universal properties of certain interacting quantum phases of matter in terms of the simpler data that characterizes the underlying layers. In the study of fracton phases, the X-Cube model in 3+1D can be realized via such a construction by starting with a stack of 2+1D Toric Codes and turning on a coupling which condenses a composite "particle-string" object. In a recent work [Phys. Rev. B 112, 125124 (2025)], we have demonstrated that in fact, the particle-string can be viewed as a symmetry defect of a topological 1-form symmetry. In this paper, we study the result of gauging this symmetry in depth. We unveil a rich gauging web relating the X-Cube model to symmetry protected topological (SPT) phases protected by a mix of subsystem and higher-form symmetries, subsystem symmetry fractionalization in the 3+1D Toric Code, and non-trivial extensions of topological symmetries by subsystem symmetries. Here, our work emphasizes the importance of topological symmetries in non-topological, geometric phases of matter.

Anyons↗

h5bench: A unified benchmark suite for evaluating HDF5 I/O performance on pre‐exascale platforms

Summary Parallel I/O is a critical technique for moving data between compute and storage subsystems of supercomputers. With massive amounts of data produced or consumed by compute nodes, high‐performant parallel I/O is essential. I/O benchmarks play an important role in this process; however, there is a scarcity of I/O benchmarks representative of current workloads on HPC systems. Toward creating representative I/O kernels from real‐world applications, we have created h5bench , a set of I/O kernels that exercise hierarchical data format version 5 (HDF5) I/O on parallel file systems in numerous dimensions. Our focus on HDF5 is due to the parallel I/O library's heavy usage in various scientific applications running on supercomputing systems. The various tests benchmarked in the h5bench suite include I/O operations (read and write), data locality (arrays of basic data types and arrays of structures), array dimensionality (one‐dimensional arrays, two‐dimensional meshes, three‐dimensional cubes), I/O modes (synchronous and asynchronous). In this paper, we present the observed performance of h5bench executed along several of these dimensions on existing supercomputers (Cori and Summit) and pre‐exascale platforms (Perlmutter, Theta, and Polaris). h5bench measurements can be used to identify performance bottlenecks and their root causes and evaluate I/O optimizations. As the I/O patterns of h5bench are diverse and capture the I/O behaviors of various HPC applications, this study will be helpful to the broader supercomputing and I/O community.

97 MATHEMATICS AND COMPUTING↗

Review of In Situ Sensing for Directed Energy Deposition for Industrial Part Quality Assessment

As the use additive manufacturing (AM) processes continues to grow in critical industries, improved quality assurance methods are becoming increasingly sought after for qualification and certification of AM components. Traditional nondestructive evaluation of printed components is often unable to supply the required confidence in print quality to justify qualification and certification, but the layer-by-layer nature of AM provides unprecedented opportunities for in situ quality inspection. This document summarizes recent developments in process monitoring research specifically related to Directed Energy Deposition (DED). Particular attention is given to three aspects of the highlighted manuscripts: (1) the type of sensors used, (2) features extracted from each sensor modality, and (3) analysis of extracted features for AM quality assessment. Based on the review of the state-of-the-art, several observations have been made. First, none of the reviewed works have applied their trained models to real part geometries, with many of the works relying on single track experiments, thin-walled structures, and cubes. Similarly, there have not been any works demonstrating model generalizability, i.e., a model trained on data from one build allows for fruitful analysis of data from another build. Many works used machine learning techniques to distinguish different process regimes (i.e., normal, keyholing, lack-of-fusion), but very few papers have investigated stochastic variation in an already “optimized” process. Sensor fusion approaches are also limited in the DED sensing literature, but the few works that have employed such techniques have demonstrated the benefits. Finally, registration of in situ data to the build coordinate system is of paramount importance to producing industrially relevant in situ monitoring systems. Data registration allows direct correlations between process anomalies detected in the process monitoring data to localized departures in part quality, but such techniques are generally lacking in the current literature.

36 MATERIALS SCIENCE↗

Flux Cube Reconstruction from Slitless Spectroscopy

Slitless spectroscopy enables efficient, large-area surveys without target preselection, yet it faces challenges from source blending, higher noise, and lost spatial–spectral information. We present an advanced, nonparametric, data-driven algorithm that leverages multiple dispersion angles to reconstruct three-dimensional flux distributions, providing low-resolution integral field unit capabilities from slitless data. By treating each pixel as an independent element, our method naturally handles source confusion without requiring prior assumptions regarding redshifts, templates, or model libraries. We validate the algorithm using simulated Roman Space Telescope wide-field slitless spectroscopy images that are equivalent to what is expected from the High-Latitude Time-Domain Survey. First, we demonstrate that a host-galaxy model reconstructed from multiple dispersion angles can be used to accurately subtract host light from a transient, recovering a Type Ia supernova spectrum with minimal bias. Second, we showcase a high-fidelity flux-cube reconstruction of a complex galaxy, successfully measuring the redshift and recovering continuum, emission, and absorption features. This approach highlights the potential of multi-dispersion-angle slitless data to provide spatially resolved spectral information in a nonparametric way, which is traditionally accessible only with integral field spectroscopy, opening a new window into large, unbiased, and spatially resolved studies of galaxy evolution.

Griggio, M. [Space Telescope Science Institute, Ba↗

Fast neutron leakage spectra of the EUCLID experiment

Special nuclear material in sub-critical and critical configurations measured in integral experiments are important for validation and adjustment of nuclear data. Many different evaluations of nuclear data exist, and these different evaluations can provide different values for individual cross sections that vary due to the uncertainties in differential experiments or lack of such data. For integral experiments, differences in these individual cross sections can have compensating errors, which lead to the same answer. One example of this is the Jezebel critical assembly, where k eff of the system is correctly computed by both ENDF/B-VIII.0 and JEFF-3.3, despite having substantially different underlying evaluated values for specific reactions (such as elastic and inelastic cross sections). To reduce compensating errors in nuclear data, the Experiments Underpinned by Computational Learning for Improvements in Nuclear Data (EUCLID) project has utilized machine learning to design a set of sub-critical and critical experiments. These experiments include slab- and cube-like configurations of 239 Pu in the form of the ZPPR plates. Six different responses were measured on a total of thirteen different configurations. One of these responses, the neutron leakage spectrum, was measured using an EJ301D detector. Finally, the results of the neutron leakage spectra show good agreement (within 1–2 σ ) with the expected spectrum from simulations and will be used in the subsequent nuclear data adjustment done by the EUCLID team.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

How to Safely Build 100-plus Kilograms of Weapons-Grade Plutonium

The goal of the EUCLID (Experiments Underpinned by Computational Learning for Improvements in Nuclear Data) project was to reduce compensating errors by utilizing machine learning to both help determine which reactions contain compensating errors as well as optimizing an experiment which can be used to maximally reduce these errors. Compensating errors can adversely impact the predictive power of application simulations, and therefore it’s useful to further constrain nuclear data and reduce these errors. The EUCLID project included building two configurations at the National Criticality Experiments Research Center (NCERC). These two configurations had very different geometries (one was cube-like and one was slab-like). Previous works focus on selection of the target experiment(s), radiation transport capabilities developed in the project, the experiment optimization, and the performance of the experiments. This work will focus only on the safety aspects of performing this experiment, which utilized over 100 kg of weapons-grade plutonium.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Studying Aerosol, Clouds, and Air Quality in the Coastal Urban Environment of Southeastern Texas

A multi-agency succession of field campaigns was conducted in southeastern Texas during July 2021 through October 2022 to study the complex interactions of aerosols, clouds and air pollution in the coastal urban environment. As part of the Tracking Aerosol Convection interactions Experiment (TRACER), the TRACER- Air Quality (TAQ) campaign the Experiment of Sea Breeze Convection, Aerosols, Precipitation and Environment (ESCAPE) and the Convective Cloud Urban Boundary Layer Experiment (CUBE), a combination of ground-based supersites and mobile laboratories, shipborne measurements and aircraft-based instrumentation were deployed. These diverse platforms collected high-resolution data to characterize the aerosol microphysics and chemistry, cloud and precipitation micro- and macro-physical properties, environmental thermodynamics and air quality-relevant constituents that are being used in follow-on analysis and modeling activities. We present the overall deployment setups, a summary of the campaign conditions and a sampling of early research results related to: (a) aerosol precursors in the urban environment, (b) influences of local meteorology on air pollution, (c) detailed observations of the sea breeze circulation, (d) retrieved supersaturation in convective updrafts, (e) characterizing the convective updraft lifecycle, (f) variability in lightning characteristics of convective storms and (g) urban influences on surface energy fluxes. The work concludes with discussion of future research activities highlighted by the TRACER model-intercomparison project to explore the representation of aerosol-convective interactions in high-resolution simulations.

54 ENVIRONMENTAL SCIENCES↗

VTK-m User's' Guide (V.1.7)

High-performance computing relies on ever finer threading. Advances in processor technology include ever greater numbers of cores, hyperthreading, accelerators with integrated blocks of cores, and special vectorized instructions, all of which require more software parallelism to achieve peak performance. Traditional visualization solutions cannot support this extreme level of concurrency. Extreme scale systems require a new programming model and a fundamental change in how we design algorithms. To address these issues we created VTK-m: the visualization toolkit for multi-/many-core architectures. VTK-m supports a number of algorithms and the ability to design further algorithms through a top-down design with an emphasis on extreme parallelism. VTK-m also provides support for finding and building links across topologies, making it possible to perform operations that determine manifold surfaces, interpolate generated values, and find adjacencies. Although VTK-m provides a simplified high-level interface for programming, its template-based code removes the overhead of abstraction. VTK-m simplifies the development of parallel scientific visualization algorithms by providing a framework of supporting functionality that allows developers to focus on visualization operations. Consider the listings in Figure 1.1 that compares the size of the implementation for the Marching Cubes algorithm in VTK-m with the equivalent reference implementation in the CUDA software development kit. Because VTK-m internally manages the parallel distribution of work and data, the VTK-m implementation is shorter and easier to maintain. Additionally, VTK-m provides data abstractions not provided by other libraries that make code written in VTK-m more versatile.This book includes contributions from the VTK-m community including the VTK-m development team and the user community.

97 MATHEMATICS AND COMPUTING↗

The VTK-m Users' Guide (V.1.9)

High-performance computing relies on ever finer threading. Advances in processor technology include ever greater numbers of cores, hyperthreading, accelerators with integrated blocks of cores, and special vectorized instructions, all of which require more software parallelism to achieve peak performance. Traditional visualization solutions cannot support this extreme level of concurrency. Extreme scale systems require a new programming model and a fundamental change in how we design algorithms. To address these issues we created VTK-m: the visualization toolkit for multi-/many-core architectures. VTK-m supports a number of algorithms and the ability to design further algorithms through a top-down design with an emphasis on extreme parallelism. VTK-m also provides support for finding and building links across topologies, making it possible to perform operations that determine manifold surfaces, interpolate generated values, and find adjacencies. Although VTK-m provides a simplified high-level interface for programming, its template-based code removes the overhead of abstraction. VTK-m simplifies the development of parallel scientific visualization algorithms by providing a framework of supporting functionality that allows developers to focus on visualization operations. Consider the listings in Figure 1.1 that compares the size of the implementation for the Marching Cubes algorithm in VTK-m with the equivalent reference implementation in the CUDA software development kit. Because VTK-m internally manages the parallel distribution of work and data, the VTK-m implementation is shorter and easier to maintain. Additionally, VTK-m provides data abstractions not provided by other libraries that make code written in VTK-m more versatile.

97 MATHEMATICS AND COMPUTING↗

The VTK-m Users' Guide (V.2.0)

High-performance computing relies on ever finer threading. Advances in processor technology include ever greater numbers of cores, hyperthreading, accelerators with integrated blocks of cores, and special vectorized instructions, all of which require more software parallelism to achieve peak performance. Traditional visualization solutions cannot support this extreme level of concurrency. Extreme scale systems require a new programming model and a fundamental change in how we design algorithms. To address these issues we created VTK-m: the visualization toolkit for multi-/many-core architectures. VTK-m supports a number of algorithms and the ability to design further algorithms through a top-down design with an emphasis on extreme parallelism. VTK-m also provides support for finding and building links across topologies, making it possible to perform operations that determine manifold surfaces, interpolate generated values, and find adjacencies. Although VTK-m provides a simplified high-level interface for programming, its template-based code removes the overhead of abstraction. VTK-m simplifies the development of parallel scientific visualization algorithms by providing a framework of supporting functionality that allows developers to focus on visualization operations. Consider the listings in Figure 1.1 that compares the size of the implementation for the Marching Cubes algorithm in VTK-m with the equivalent reference implementation in the CUDA software development kit. Because VTK-m internally manages the parallel distribution of work and data, the VTK-m implementation is shorter and easier to maintain. Additionally, VTK-m provides data abstractions not provided by other libraries that make code written in VTK-m more versatile.

97 MATHEMATICS AND COMPUTING↗

Chemical signature characterization with hyperspectral imagery: novel deep learning model architectures and physically-motivated data augmentation techniques

The high spectral resolution afforded by Hyperspectral Imaging (HSI) sensors is poised to bring unprecedented advancements to signature characterization applications. Thus far, much of the research in the machine learning field devoted to HSI applications has focused on a few specific tasks like land-use land-cover classification. In land classification tasks, spatial information is very important, and model architectures are often designed to leverage spatial contexts. However, it is unclear how well these spatially-tuned models will translate to tasks where spectral information is critical, like the detection and characterization of chemicals. In this work, we compare spectral models (inputs are 1D spectra) and spatial-spectral models (inputs are 3D cubes) in the context of predicting chemical concentration maps. We find that spatial-spectral models perform the best, though we find a wide range in performance across the different architectures tested. Additionally, we find that model performance is impacted by the availability of training data, particularly in scenarios where the training data doesn't fully capture the true variance of real-world conditions. We find that data augmentation can help mitigate sparse coverage of observed parameter space (e.g., seasonal or geographic variability in ground cover), and present augmentation strategies that are tailored to hyperspectral data.

• Artificial intelligence (AI) / machine learning ↗

ILAMBv2.7 benchmarking results comparing E3SMv2.1 land-atmosphere coupled (BGCv2LNDATM) and stand alone land (ELM) simulations with CMIP6 emission driven historical simulations

This dataset contains land model benchmarking results for the Energy Exascale Earth System Model version 2.1 (E3SMv2.1), including outputs from both coupled biogeochemistry simulations and stand-alone land model simulations. These results are compared against several emission-driven historical simulations from the Coupled Model Intercomparison Project Phase 6 (CMIP6). Benchmarking was conducted using the International Land Model Benchmarking (ILAMB) package, version 2.7 (ILAMBv2.7). CMIP6 model outputs were sourced from the Earth System Grid Federation (ESGF), while the E3SMv2.1 results were derived from raw model outputs. These outputs underwent processing steps such as time serialization, conservative regridding, and data standardization to ensure comparability. For spatial interpolation, the Earth System Modeling Framework (ESMF) tool, ESMF_RegridWeightGen, was employed to generate regridding weights, enabling the transformation of E3SM’s native cubed-sphere grid to a regular latitude-longitude grid.

Feng, Sha [PNNL]↗

Reaction Rate Ratios for Recent Fast Metal Experiments with Large Plutonium Masses

Reaction rate ratios are integral responses that are used within the criticality experiments field because they contain spectral information. While these types of measurements have been utilized for nuclear data validation with historic experiments, few experiments of this type have been utilized for recent experiments, as few exist. This work focuses on measured reaction rate ratios for two nearly bare plutonium critical assemblies with different geometries: one that is cube like (with a Pu mass of 40 kg) and one that is slab like (with a Pu mass of 109 kg). Irradiations were performed with both configurations in which foils were placed near the center of the assembly. Plutonium, highly enriched uranium, depleted uranium, and Au foils were included in the irradiation and counted via high-purity germanium detectors. From these measurements, reaction rate ratios were calculated. Measured and simulated values and uncertainties are presented for the reaction rate ratios. Ratios utilizing the following reactions are given in this work: 197 Au(n, γ), 197 Au(n,2n), 235 U(n,fission), 238 U(n, fission), 238 U(n,2n), 238 U(n,γ), and 239 Pu(n,fission). Uncertainties for the measured reaction rate ratios ranged from 4% to 7%, and the contribution of various parameters to this uncertainty was investigated. The results are compared to historical experiments and should be used for nuclear data validation for future nuclear data library releases. These measurements are part of the EUCLID (Experiments Underpinned by Computational Learning for Improvements in Nuclear Data) project, which utilizes measurement responses in addition to k eff (such as these reaction rate ratios) to help reduce uncertainties in 239 Pu nuclear data.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗