Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “standardized data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 163 records · Page 9

Carbonate clumped isotope analysis (Δ 47 ) of 21 carbonate standards determined via gas‐source isotope‐ratio mass spectrometry on four instrumental configurations using carbonate‐based standardization and multiyear data sets

Rationale Clumped isotope geochemistry examines the pairing or clumping of heavy isotopes in molecules and provides information about the thermodynamic and kinetic controls on their formation. The first clumped isotope measurements of carbonate minerals were first published 15 years ago, and since then, interlaboratory offsets have been observed, and laboratory and community practices for measurement, data analysis, and instrumentation have evolved. Here we briefly review historical and recent developments for measurements, share Tripati Lab practices for four different instrument configurations, test a recently published proposal for carbonate‐based standardization on multiple instruments using multi‐year data sets, and report values for 21 different carbonate standards that allow for recalculations of previously published data sets. Methods We examine data from 4628 standard measurements on Thermo MAT 253 and Nu Perspective IS mass spectrometers, using a common acid bath (90°C) and small‐sample (70°C) individual reaction vessels. Each configuration was investigated by treating some standards as anchors (working standards) and the remainder as unknowns (consistency standards). Results We show that different acid digestion systems and mass spectrometer models yield indistinguishable results when instrument drift is well characterized. For linearity correction, mixed gas‐and‐carbonate standardization or carbonate‐only standardization yields similar results. No difference is observed in the use of three or eight working standards for the construction of transfer functions. Conclusions We show that all configurations yield similar results if instrument drift is robustly characterized and validate a recent proposal for carbonate‐based standardization using large multiyear data sets. Δ 47 values are reported for 21 carbonate standards on both the absolute reference frame (ARF; also refered to as the Carbon Dioxide Equilibrated Scale or CDES) and the new InterCarb‐Carbon Dioxide Equilibrium Scale (I‐CDES) reference frame, facilitating intercomparison of data from a diversity of labs and instrument configurations and restandardization of a broad range of sample sets between 2006, when the first carbonate measurements were published, and the present.

58 GEOSCIENCES↗

Perspectives for artificial intelligence in bioprocess automation

Recent advances in artificial intelligence (AI) have rapidly changed the lab automation landscape, promoting self-driving laboratories (SDLs) that enable autonomous scientific discovery. These trends are increasingly applied in bioprocess development, yet bioprocessing faces unique challenges - biological complexity, regulatory and safety requirements, and multiscale experimentation - that distinguish it from other automation domains. Rather than pursuing full autonomy, we foresee that hybrid SDLs, combining AI-driven decision-making with sustained human oversight, represent the most practical near-term trajectory. This review examines three interconnected perspectives: (i) hybrid human-machine decision-making for bioprocessing; (ii) laboratory design considerations in the era of AI; and (iii) scale-up challenges when transitioning from screening to manufacturing. We highlight critical gaps in data standardization and the required community efforts necessary to realize autonomous bioprocess innovation.

Helleckes, Laura Marie↗

A systematic literature review on residential demand response with a focus on opportunities for low-income communities

Demand Response improves efficiency and grid stability, reducing peak load and total system costs. It can potentially contribute to other societal aspects, including mitigating climate change and improving air quality, health, stress levels, and comfort. This paper reviewed the literature on residential demand response programs in the United States, through the lens of the five energy justice principles, identifying and categorizing a set of 165 papers from 1960 to 2024. As a result, we observe significant progress in distributive justice, however, mostly focused on increasing participation rather than attending specific needs of historically marginalized communities. Concordantly, recognition and procedural justice dimensions are lagging, while cosmopolitan justice studies are scarce, underscoring a need to improve the measurement of impacts resulting from the implementation of demand response programs. Future research opportunities include the consolidation of standardized definitions and metrics, to better understand burdens and needs of diverse residential users. In terms of policy, collection and transparent access to standardized data is required as a condition to the design of flexible, customizable, transparent, and easy to navigate demand response programs. As a result, a better understanding of and engagement with customers has the potential to increase adoption rates, persistence in the program, and overall impacts for the system, households, and society.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Data driven investigation to understand the influence of total solids on biological biogas upgrading

In situ biogas upgrading achieves CO 2 conversion to CH 4 via hydrogenotrophic methanogenesis; however, gas-liquid mass transfer constraints limit the upgrading performance. Recognizing that optimization studies often underrepresent the effects of total solids (TS) and organic loading rate (OLR), this study undertook a holistic, statistics driven assessment of operating conditions for in situ H 2 assisted biogas upgrading, centering the analysis on TS and OLR. A dataset of 31 studies was compiled and comprised 99 observations. A rigorous analytical framework was employed, combining data standardization, fixed- and random-effects (REML) weighted regressions with cluster-robust errors, stratified analyses, and machine learning. Mixed-effects meta regression indicated that TS was the main factor explaining differences of methane fraction (CH 4 %) when considering the between studies heterogeneity. Focusing on a near-stoichiometric subset (H 2 /CO 2 ≈ 4:1), TS remained significant. Stratified results showed a stronger negative relationship between TS and CH 4 % in UASB reactors than in CSTRs, with a negative effect under mesophilic conditions and no significant effect under thermophilic conditions. A Random Forest model corroborated the statistical findings, consistently ranking H 2 /CO 2 ratio, OLR, TS, and hydrogen injection rate (HIR) as the most influential predictors. These findings delineate trends across increasing TS levels, particularly between 1% and 10%, and provide preliminary insights for TS above 15% in in situ biogas upgrading. They further provide insights for the influence of TS by reactor type and temperature, thereby advancing the evidence base for implementing biological CO 2 conversion to CH 4 in practice.

In situ biogas upgrading↗

Large language models for batteries

Large Language Models (LLMs) are advanced artificial intelligence systems capable of solving diverse tasks using language, reasoning, and external tools. Despite their growing deployment in academia and industry, their potential remains underexplored in battery research. This review presents a comprehensive overview of existing and emerging applications of LLMs in batterie field, addressing two critical questions: What can LLMs offer to support battery-related tasks, and how to develop more effective models for this purpose. We begin by outlining the principles of LLMs and criteria for selecting appropriate models and tools for battery research and development. We then explore their roles in text-mining, data interpretation, and the development of intelligent battery systems. In parallel, we discuss technical challenges, such as data standardizing and sharing, model evaluation, and tool integration. Lastly, we propose future research directions with short-, medium-, and long-term goals and highlight more broad perspectives for connecting experts and cross-disciplinary collaborations.

SoC↗

A Roadmap for Edge Computing Enabled Automated Multidimensional Transmission Electron Microscopy

The advent of modern, high-speed electron detectors has made the collection of multidimensional hyperspectral transmission electron microscopy datasets, such as 4D-STEM, a routine. However, many microscopists find such experiments daunting since analysis, collection, long-term storage, and networking of such datasets remain challenging. Some common issues are their large and unwieldy size that often are several gigabytes, non-standardized data analysis routines, and a lack of clarity about the computing and network resources needed to utilize the electron microscope. The existing computing and networking bottlenecks introduce significant penalties in each step of these experiments, and thus, real-time analysis-driven automated experimentation for multidimensional TEM is challenging. One solution is to integrate microscopy with edge computing, where moderately powerful computational hardware performs the preliminary analysis before handing off the heavier computation to high-performance computing (HPC) systems. In this work, we trace the roots of computation in modern electron microscopy, demonstrate deep learning experiments running on an edge system, and discuss the networking requirements for tying together microscopes, edge computers, and HPC systems.

47 OTHER INSTRUMENTATION↗

Experimental and Computational Evaluation of Lipidomic In-Source Fragmentation as a Result of Postionization with Matrix-Assisted Laser Desorption/Ionization

Matrix-assisted laser desorption/ionization mass spectrometry imaging (MALDI-MSI) can provide spatially resolved molecular information about a sample. Recently, a postionization approach (MALDI-2) has been commercially integrated with MALDI-MSI, allowing for bettered sensitivity and consequent improved spatial resolution. While advantages of MALDI-2 have previously been established, we demonstrate here statistically increased in-source fragmentation (ISF) results from postionization with a commercial instrument. Via lipid standard analyses, known MALDI ISF pathways (e.g., loss of trimethylamine) were statistically increased in MALDI-2 compared to MALDI-1 (65–172% increase in fragmentation). Gas phase molecular modeling with density functional theory estimated that the most-weighted virtual orbitals to excite within lipids involve ester and phosphate bonds. Protonated lipid excitation energies are furthermore red-shifted compared to those of other adduct types [e.g., 254 nm for protonated PC(16:0/18:1)] and approach the MALDI-2 laser energy (266 nm). Analysis of rat brain homogenate detected statistically more positive-ion mode peaks with MALDI-2 (1090) than that with MALDI-1 (719), where Kernel density estimations showed that the majority of this enhancement occurs with low m/z ions (i.e., m/z 75–500). Taken together with the lipid standard data, these observations may indicate ISF due to postionization. Finally, while artifact contributions from matrix blanks were also noted, both experimental and computational data sets suggest that the overall extent of ISF is statistically increased in MALDI-2 compared to MALDI-1.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Large field-of-view event-mode camera for high-precision epithermal neutron resonance imaging

A large-area event-mode camera system coupled with a 6 LiF-ZnS:Ag scintillator is applied for neutron resonance imaging (NRI) on the energy-resolved neutron imaging (ERNI) flight path, also known as Flight Path 5 (FP5), at the Los Alamos Neutron Science Center (LANSCE). This novel neutron imaging system, featuring a 120 x 120 mm 2 field of view, efficiently captures resonance information across the entire image in a single acquisition, significantly reducing beam time requirements compared to conventional energy-resolved neutron imaging systems. High-quality neutron radiographs with enhanced spatial resolution are achieved through the reconstruction of neutron events based on observations of individual photons emitted from the scintillator. The system demonstrates reduced background through neutron/gamma discrimination capabilities while maintaining sharpness across a large fields of view. In the measurements presented here, a spatial resolution of approximately 340 μm was achieved using center-of-gravity photon cluster centroiding. We demonstrate the system’s capability for quantitatively determining isotopic distributions in various thin samples, as well as automatically reconstructing complex scenes with overlapping resonances from diverse samples. These results are obtained using standard data analysis tools, despite the relatively slow 6 LiF-ZnS:Ag scintillator, which may not be optimal for absorption resonance detection. The capabilities demonstrated here offer a valuable, versatile, and cost-effective solution for high spatial and temporal resolution, large field-of-view energy-resolved neutron imaging, with potential applications across various scientific and industrial domains.

36 MATERIALS SCIENCE↗

Ptychography at all wavelengths

Ptychography is a computational imaging technique that operates across multiple wavelength regimes, from electron (picometres) to X-ray (~0.1 nm), extreme ultraviolet (~10 nm) and visible light (micrometres). By reconstructing both amplitude and phase from diffraction patterns, ptychography enables high-resolution, quantitative imaging without conventional limitations imposed by lens-based optics. Ptychography has enabled advances across a range of scales: achieving deep-sub-angstrom resolution with electron microscopy, becoming an indispensable tool at X-ray synchrotron facilities worldwide and overcoming the trade-offs between resolution and field-of-view in optical imaging. This Primer provides a unified treatment of ptychography across these wavelength regimes. First, we discuss theoretical foundations, reconstruction algorithms, experimental considerations and wavelength-specific challenges. We then give examples of raw and processed data from various configurations and wavelengths. Next, we highlight key applications of ptychography in life sciences, materials science and industry. We also discuss data standards, open-source software implementations and best practices for ensuring reproducibility across different wavelength regimes. Finally, we consider limitations and future opportunities for ptychography. Together with accompanying datasets and code implementations, this Primer aims to serve newcomers and experienced practitioners in the field, facilitating broader adoption of ptychography across different disciplines.

47 OTHER INSTRUMENTATION↗

Unifying thermochemistry concepts in computational heterogeneous catalysis

Thermophysical properties of adsorbates and gas-phase species define the free energy landscape of heterogeneously catalyzed processes and are pivotal for an atomistic understanding of the catalyst performance. These thermophysical properties, such as the free energy or the enthalpy, are typically derived from density functional theory (DFT) calculations. Enthalpies are species-interdependent properties that are only meaningful when referenced to other species. The widespread use of DFT has led to a proliferation of new energetic data in the literature and databases. However, there is a lack of consistency in how DFT data is referenced and how the associated enthalpies or free energies are stored and reported, leading to challenges in reproducing or utilizing the results of prior work. Additionally, DFT suffers from exchange–correlation errors that often require corrections to align the data with other global thermochemical networks, which are not always clearly documented or explained. In this review, we introduce a set of consistent terminology and definitions, review existing approaches, and unify the techniques using the framework of linear algebra. This set of terminology and tools facilitates the correction and alignment of energies between different data formats and sources, promoting the sharing and reuse of ab initio data. Standardization of thermochemistry concepts in computational heterogeneous catalysis reduces computational cost and enhances fundamental understanding of catalytic processes, which will accelerate the computational design of optimally performing catalysts.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

A customizable data management framework for high-repetition-rate high-energy-density science

The high-energy-density (HED) physics community is moving toward a new paradigm of high-repetition-rate (HRR) operation. To fully leverage the scientific power of HRR HED facilities, all of the components of each subsystem (laser, targetry, and performance diagnostics) must be connected and synchronized in a reliable and robust manner while the data acquired are tagged and archived in real time. To this end, GA has begun developing a generalized NoSQL-database framework, the MongoDB repository for information and archiving. An organizational strategy has been developed that shifts HED data organization from a shot-based to a diagnostic-based approach in order to increase archival and retrieval efficiency that lends itself to optimization applications. This work is a first step in pushing HRR HED science toward data management solutions that emphasize machine actionability and aim to stimulate community engagement to define data standards in HED science.

Instruments & Instrumentation↗

A change language for ontologies and knowledge graphs

Ontologies and knowledge graphs (KGs) are general-purpose computable representations of some domain, such as human anatomy, and are frequently a crucial part of modern information systems. Most of these structures change over time, incorporating new knowledge or information that was previously missing. Managing these changes is a challenge, both in terms of communicating changes to users and providing mechanisms to make it easier for multiple stakeholders to contribute. To fill that need, we have created KGCL, the Knowledge Graph Change Language (https://github.com/INCATools/kgcl), a standard data model for describing changes to KGs and ontologies at a high level, and an accompanying human-readable Controlled Natural Language (CNL). This language serves two purposes: a curator can use it to request desired changes, and it can also be used to describe changes that have already happened, corresponding to the concepts of “apply patch” and “diff” commonly used for managing changes in text documents and computer programs. Another key feature of KGCL is that descriptions are at a high enough level to be useful and understood by a variety of stakeholders—e.g. ontology edits can be specified by commands like “add synonym ‘arm’ to ‘forelimb’” or “move ‘Parkinson disease’ under ‘neurodegenerative disease’.” We have also built a suite of tools for managing ontology changes. These include an automated agent that integrates with and monitors GitHub ontology repositories and applies any requested changes and a new component in the BioPortal ontology resource that allows users to make change requests directly from within the BioPortal user interface. Overall, the KGCL data model, its CNL, and associated tooling allow for easier management and processing of changes associated with the development of ontologies and KGs.

96 KNOWLEDGE MANAGEMENT AND PRESERVATION↗

The Case for Strong Scaling in Deep Learning: Training Large 3D CNNs with Hybrid Parallelism

Here, we present scalable hybrid-parallel algorithms for training large-scale 3D convolutional neural networks. Deep learning-based emerging scientific workflows often require model training with large, high-dimensional samples, which can make training much more costly and even infeasible due to excessive memory usage. We solve these challenges by extensively applying hybrid parallelism throughout the end-to-end training pipeline, including both computations and I/O. Our hybrid-parallel algorithm extends the standard data parallelism with spatial parallelism, which partitions a single sample in the spatial domain, realizing strong scaling beyond the mini-batch dimension with a larger aggregated memory capacity. We evaluate our proposed training algorithms with two challenging 3D CNNs, CosmoFlow and 3D U-Net. Our comprehensive performance studies show that good weak and strong scaling can be achieved for both networks using up to 2K GPUs. More importantly, we enable training of CosmoFlow with much larger samples than previously possible, realizing an order-of-magnitude improvement in prediction accuracy.

97 MATHEMATICS AND COMPUTING↗

Characterization and differentiation of aluminum powders used in improvised explosive devices. Part 2: Micromorphometric method refinement and preliminary statistical analysis

Abstract Aluminum (Al) powder is commonly encountered in improvised explosive devices (IEDs) as a metallic fuel due to its availability and low cost. Although available commercially in powder form, amateur bomb‐makers also produce their own Al powder via simple methods found online. In order to provide investigative leads and forensic intelligence, it is important to evaluate not only the composition of homemade devices, but also to distinguish between the various forms of Al powder they contain. To achieve this goal, a method using automated microscopy in combination with statistical techniques has been demonstrated to have the potential to provide source discrimination and investigative leads in source attribution of Al powders in IEDs. The present research refined this method and investigated 59 industrially and amateurly produced Al powder sources with seven subsamples per source using two traditional linear discriminant analyses (LDA), one with a standard data split for training and testing, and another using leave‐one‐out cross‐validation. Averaging the classification accuracies for the two LDA‐based analyses, LDA has the ability to correctly classify 59.26%, 83.35%, and 80.69% of the samples based on their powder source, type, and production method, respectively. This classification accuracy represents a 3407%, 317%, and 61.38% increase in accuracy from random class assignment, respectively. Further, in most instances of incorrect data attribution to a particular source, the subsample has been misidentified with another sample of the same powder type or production method.

Ommen, Danica M.↗

Long-Term Stability of Mullite and Magnesia-encased Ag|Ag + Reference Electrodes in Molten MgCl 2 -KCl-NaCl

In this study, the stability of Ag|Ag + molten chloride reference electrodes was studied using mullite and magnesia tubes as reference electrode membrane materials in molten MgCl 2 -KCl-NaCl for continuous immersion up to 31 d. Two electrochemical methods were used to characterize the long-term performance of the reference electrodes at 500 °C in an inert atmosphere using a Mg rod as the working electrode: open circuit potential (OCP) and onset of Mg(II) reduction as measured by cyclic voltammetry (CV). There was only a 3% difference between the average OCP measured versus the mullite and MgO reference electrodes. Potential drift was higher for mullite (−1.34 mV d −1 ) than MgO (0.22 mV d −1 ). But both drift rates were very low compared to the OCP data standard deviation, indicating that no significant drift was measured for us to 31 d.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

AqNWB v0.1

This is a C ++ acquisition API for neurophysiological data in the NeurodataWithoutBorders (NWB) data standard.

Prince, Stephanie↗

Common Electric Power Transmission System Model JSON Schema Specification

The Common Electric Power Transmission System Model (CTM) is an intuitive, extensible, language-agnostic, and error-resistant specification of electric power network components parameter names and units, and relation between components, intended for use by the research community developing new computational methods for power systems operations and simulation. Power system datasets following the CTM specification can be read as dictionaries and manipulated in that form in most programming languages (e.g., Python, Julia, C++). This standard data structure in CTM makes it easy to work in multiple power systems domains (e.g., economic operation, reliability assessment, electricity markets, stability assessment, etc.) without requiring conversions between use-case-specific file formats with information loss in the process. This repository specifies CTM as a JSON Schema, provides documentation, derivate (code-generated) implementations of CTM, and example data and usage of the schema for important use cases.

Aravena Solis, Ignacio↗