Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Information Automation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 163 records · Page 9

Can We Fix It Automatically? Development of Fault Auto-Correction Algorithms for HVAC and Lighting Systems

A fault detection and diagnostics (FDD) tool is a type of energy management and information system designed to continuously identify the presence of faults and efficiency improvement opportunities through a one-way interface to the building automation system and application of automated analytics. Building owners and operators at the leading edge of technology adoption are using FDD tools to enable average whole-building portfolio savings of 8 percent. Although FDD tools can inform building operators of operational faults, currently a manual action is always required to correct faults and generate the associated energy savings. A subset of faults, however, such as biased sensors and manual override, can be addressed automatically, removing the need for operations and maintenance staff intervention. Automating this fault “correction” can significantly increase the savings generated by FDD tools and reduce the reliance on human intervention. Doing so is expected to advance the usability, as well as the technical and economic performance, of FDD technologies. In this paper, we present the development of 10 innovative fault auto-correction algorithms for HVAC and lighting systems. When the auto-correction routine is triggered, it will overwrite the control setpoints or other variables (via BACnet or other protocol) to implement the intended changes. These algorithms are able to automatically correct the faults or improve the operation associated with an incorrectly programmed schedule, override manual control, sensor bias, control hunting, rogue zone, and less aggressive setpoints/setpoints setback. The paper will also discuss the implementation of the auto-correction algorithms in FDD software products.

Lin, Guanjing↗

A review on recent machine learning applications for imaging mass spectrometry studies

Imaging mass spectrometry (IMS) is a powerful analytical technique widely used in biology, chemistry, and materials science fields that continue to expand. IMS provides a qualitative compositional analysis and spatial mapping with high chemical specificity. The spatial mapping information can be 2D or 3D depending on the analysis technique employed. Due to the combination of complex mass spectra coupled with spatial information, large high-dimensional datasets (hyperspectral) are often produced. Therefore, the use of automated computational methods for an exploratory analysis is highly beneficial. The fast-paced development of artificial intelligence (AI) and machine learning (ML) tools has received significant attention in recent years. These tools, in principle, can enable the unification of data collection and analysis into a single pipeline to make sampling and analysis decisions on the go. There are various ML approaches that have been applied to IMS data over the last decade. Here, in this review, we discuss recent examples of the common unsupervised (principal component analysis, non-negative matrix factorization, k-means clustering, uniform manifold approximation and projection), supervised (random forest, logistic regression, XGboost, support vector machine), and other methods applied to various IMS datasets in the past five years. The information from this review will be useful for specialists from both IMS and ML fields since it summarizes current and representative studies of computational ML-based exploratory methods for IMS.

47 OTHER INSTRUMENTATION↗

From Text to Maps: LLM-Driven Extraction and Geotagging of Epidemiological Data

Epidemiological datasets are essential for public health analysis and decision-making, yet they remain scarce and often difficult to compile due to inconsistent data formats, language barriers, and evolving political boundaries. Traditional methods of creating such datasets involve extensive manual effort and are prone to errors in accurate location extraction. To address these challenges, we propose utilizing large language models (LLMs) to automate the extraction and geotagging of epidemiological data from textual documents. Our approach significantly reduces the manual effort required, limiting human intervention to validating a subset of records against text snippets and verifying the geotagging reasoning, as opposed to reviewing multiple entire documents manually to extract, clean, and geotag. Additionally, the LLMs identify information often overlooked by human annotators, further enhancing the dataset’s completeness. Our findings demonstrate that LLMs can be effectively used to semi-automate the extraction and geotagging of epidemiological data, offering several key advantages: (1) comprehensive information extraction with minimal risk of missing critical details; (2) minimal human intervention; (3) higher-resolution data with more precise geotagging; and (4) significantly reduced resource demands compared to traditional methods.

Harrod, Karly↗

Development of Microreactor Automated Control System (MACS): Surrogate Plant-level Modeling and Control Algorithms Integration

This report discusses progress on the modeling and integration task as part of the development of Microreactor Automated Control System (MACS). What follows is a discussion of the software tools used to develop a surrogate microreactor plant-level model, MACS module development, and a summary of the findings from integration with a hardware-in-the-loop (HIL) setup at Idaho National Laboratory (INL). Oak Ridge National Laboratory worked with INL to understand available hardware at INL (such as existing control platforms). With this information and using a preliminary framework for MACS with software interfaces defined, surrogate models and initial automated control algorithms were created for integration into the hardware to provide a HIL demonstration platform for MACS. Available reduced order reactor models that leverage existing microreactor neutronics and thermal hydraulics models have been integrated in the surrogate plant-level model. The surrogate model has been evaluated for sensitivity of all parameters and the MACS software modules (including the surrogate model) have been demonstrated to interact with the hardware. The integration, however, pointed to the need for further improvements to the framework—and specifically improvements to the interfaces to ensure that the HIL simulator is capable of longer-term stable operation with MACS in the loop. These improvements will be the focus of future research.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Architectural Approaches for Integrating ADMS and DERMS: Challenges, Comparisons, and Real-World Use Cases

The electrical distribution landscape is rapidly transforming due to the proliferation of distributed energy resources (DERs) such as solar panels, wind turbines, battery storage systems, combined heat and power units, and electric vehicles, introducing variability and uncontrollability that traditional grid operators are ill-equipped to manage. This transformation is further accelerated by advancements in Information and Communication Technology infrastructure that connects control centers with end devices, demanding automation and a deeper understanding of new technologies by utility personnel. Advanced grid control techniques using system-level optimization, Artificial Intelligence, and Machine Learning at the enterprise level and distributed level are evolving to address these issues. There is also an opportunity to utilize the enormous data created by these new DER technologies in the grid. Advanced Distribution Management Systems (ADMS) and Distributed Energy Resource Management Systems (DERMS) are critical in addressing these challenges by automating grid operations and enhancing reliability. Given the relatively recent development of ADMS and DERMS, and the still relatively low level of ADMS and DERMS deployment in the industry, there is a notable deficiency in the comprehensive understanding of the challenges and benefits associated with these new technologies, especially with their complementary natures and integration architectures. This paper aims to bridge the knowledge gap in ADMS and DERMS integration, presenting three distinct integration architectures currently available, and discussing the challenges and benefits of each architecture to guide utilities, industry professionals, and researchers in optimizing grid management and decision-making processes for a resilient and efficient energy future.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Toward fully automated UED operation using two-stage machine learning model

To demonstrate the feasibility of automating UED operation and diagnosing the machine performance in real time, a two-stage machine learning (ML) model based on self-consistent start-to-end simulations has been implemented. This model will not only provide the machine parameters with adequate precision, toward the full automation of the UED instrument, but also make real-time electron beam information available as single-shot nondestructive diagnostics. Furthermore, based on a deep understanding of the root connection between the electron beam properties and the features of Bragg-diffraction patterns, we have applied the hidden symmetry as model constraints, successfully improving the accuracy of energy spread prediction by a factor of five and making the beam divergence prediction two times faster. The capability enabled by the global optimization via ML provides us with better opportunities for discoveries using near-parallel, bright, and ultrafast electron beams for single-shot imaging. It also enables directly visualizing the dynamics of defects and nanostructured materials, which is impossible using present electron-beam technologies.

36 MATERIALS SCIENCE↗

When ChatGPT Meets Vulnerability Management: The Good, the Bad, and the Ugly

Vulnerability management is a very challenging and time-consuming task. For many organizations, security operators need to learn about the properties of vulnerabilities to prioritize and mitigate them. Due to the lack of automated tools for vulnerability assessment, operators usually manually search for and read related information from sources online. Recent advances in large language models, like ChatGPT, open up an opportunity for time savings and may prompt operators to use these models as vulnerability information sources. In this work, we evaluate the ability of ChatGPT and several of its siblings to accurately answer user questions about vulnerability properties as well as to provide information for how to mitigate a vulnerability. We also explore their summarization capabilities when multiple vulnerability advisory documents are provided. We find that the models perform poorly on information retrieval tasks, but they perform quite well on summarization.

McClanahan, Kylie↗

MADA: Multi-Agent Design Assistant

MADA (Multi-Agent Design Assistant) is a Large Language Model (LLM) powered multi-agent framework that coordinates specialized agents for complex design workflows. The system was designed for HPC workflows with the following agents in mind: 1) A Job Management Agent (JMA) launches and manages ensemble simulations on HPC systems, 2) a Geometry Agent (GA) generates meshes, and 3) an Inverse Design Agent (IDA) proposes new designs informed by simulation outcomes. Our framework reduces cumbersome manual workflow setup, and enables automated design exploration at scale. However, the software also enables users to rapidly create new multi-agent systems. Simply define new agents in a configuration file, giving each their own set of tools (via MCP), and then chat and prompt your new multi-agent system. Is

Gunnarson, BrianS [Lawrence Livermore National Lab↗

Boosting Energy Efficiency of Heterogeneous Connected Automated Vehicle (CAV) Fleets via Anticipative and Cooperative Vehicle Guidance

In 2017, the Department of Energy funded a team at Clemson University and Argonne National Laboratory to develop collaborative perception and anticipative/predictive vehicle guidance schemes for Connected and Automated Vehicles (CAVs) and to quantify the energy saving potential of this technology in large scale traffic microsimulations at different levels of technology penetration and also experimentally. The project goal was demonstrating up to a 10% energy saving potential from different aspects of the implementation with a focus on reducing unnecessary braking events by anticipatory speed and lane selection. This project introduced novel anticipative car following and lane selection schemes for Connected and Automated Vehicles (CAVs). Our control schemes benefited from prediction of human driver behavior, information exchange between CAVs, and sometimes from collaboration to save energy, reduce braking events, and harmonize traffic. The energy savings was first demonstrated by traffic micro-simulations and then via a novel Vehicle-In-the-Loop (VIL) experimental testbed.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Evaluating the Effectiveness of a Detection and Deterrent System in Reducing Golden Eagle Fatalities at Operational Wind Facilities

The Renewable Energy Wildlife Institute (REWI) was appointed as the prime awardee of DOE award number DE-EE0007883 to lead a team of scientists, wind developers, and technology manufacturers toward the overarching goal of evaluating the effectiveness of the current DTBird system in minimizing the risk of golden eagles (Aquila chrysaetos) and other large soaring raptors from approaching the rotor-swept zone (RSZ) of operating wind turbines. As part of this goal, the team set out to 1) quantify the expected reduction in collision risk for golden eagles from operation of the detection and deterrence modules in a manner that supports the approach used by the U.S. Fish and Wildlife Service (USFWS) to assess and credit facility operators for their efforts to minimize predicted collision fatalities and 2) provide information to help improve the technology to maximize its effectiveness. DTBird is an automated detection and audio deterrent system created by the Spanish company Liquen, designed to discourage birds from entering the RSZ of spinning wind turbines. The system uses cameras to automatically detect airborne targets of interest, records each such event in an online database, and triggers a warning signal (loud sound) if the tracked object has moved close to the turbine. If the object moves even closer to the RSZ, a more aggressive dissuasion signal is broadcast. To meet our objectives, the team conducted a two-year experiment at the Goodnoe Hills wind facility in Washington state, in which 14 turbines were outfitted with DTBird units. Daily, each DTBird-equipped turbine was randomly assigned to a control or treatment group. Treatment turbines operated with DTBird running as intended—broadcasting warning or deterrent signals when DTBird detected a target within range. On control turbines, no sound signals were broadcast if a moving target triggered the DTBird system. The team also flew unmanned aerial vehicles (UAVs) designed to coarsely mimic the general size, weight, and coloration of golden eagles in programmed flight transects across DTBird detection ranges to quantify DTBird’s ability to detect intended targets and to evaluate factors that influence the probability of detection and DTBird’s response distances. Additionally, the team evaluated the behavioral responses of in situ eagles exposed to spinning turbines alone (visual and sound influences) versus spinning turbines plus broadcasted DTBird audio deterrents, to estimate the effectiveness of deterrence by the DTBird system. The data and results from these investigations were combined with those from a pilot study conducted at the Manzana Wind Power Project in California to better evaluate DTBird’s effectiveness across different landscapes.

17 WIND ENERGY↗

RivGraph: Automatic extraction and analysis of river and delta channel network topology

River networks sustain life and landscapes by carrying and distributing water, sediment, and nutrients throughout ecosystems and communities. At the largest scale, river networks drain continents through tree-like tributary networks. At typically smaller scales, river deltas and braided rivers form loopy, complex distributary river networks via avulsions and bifurcations.In order to model flows through these networks or analyze network structure, the topology, or connectivity, of the network must be resolved. Additionally, morphologic properties of each river channel as well as the direction of flow through the channel inform how fluxes travel through the network’s channels. Riv Graphis a Python package that automates the extraction and characterization of river channel networks from a user-provided binary image, or mask, of a channel network (Fig. 1). Masks may be derived from (typically remotely-sensed) imagery, simulations, or even hand-drawn. RivGraph will create explicit representations of the channel network by resolving river centerlines as links, and junctions as nodes. Flow directions are solved for each link of the network without using auxiliary data, e.g., a digital elevation model (DEM). Morphologic properties are computed as well, including link lengths, widths, sinuosities, branching angles,and braiding indices. If provided,RivGraph will preserve georeferencing information of the mask and will export results as ESRI shapefiles, GeoJSONs, and GeoTIFFs for easy import into GIS software.RivGraph can also return extracted networks as networkx objects for convenient interfacing with the full-featured networkx package (Hagberg et al., 2008). Finally, RivGraph offers a suite of topologic metrics that were specifically designed for river channel network analysis (Tejedor et al., 2015b).

54 ENVIRONMENTAL SCIENCES↗

Coexistence and Interplay of Two Ferroelectric Mechanisms in Zn 1-x Mg x O

Ferroelectric materials promise exceptional attributes including low power dissipation, fast operational speeds, enhanced endurance, and superior retention to revolutionize information technology. However, the practical application of ferroelectric-semiconductor memory devices has been significantly challenged by the incompatibility of traditional perovskite oxide ferroelectrics with metal-oxide-semiconductor technology. Recent discoveries of ferroelectricity in binary oxides such as Zn 1-x Mg x O and Hf 1-x Zr x O have been a focal point of research in ferroelectric information technology. Here, this work investigates the ferroelectric properties of Zn 1-x Mg x O utilizing automated band excitation piezoresponse force microscopy. This findings reveal the coexistence of two ferroelectric subsystems within Zn 1-x Mg x O. A “fringing-ridge mechanism” of polarization switching is proposed that is characterized by initial lateral expansion of nucleation without significant propagation in depth, contradicting the conventional domain growth process observed in ferroelectrics. This unique polarization dynamics in Zn 1-x Mg x O suggests a new understanding of ferroelectric behavior, contributing to both the fundamental science of ferroelectrics and their application in information technology.

36 MATERIALS SCIENCE↗

Comparison of automated chemical-guided segmentation and human annotation of soil organic matter in X-ray microcomputed tomography imaging in contrasted soil types

Soil organic matter (OM) formation and persistence is strongly influenced by the spatial distribution of organic substrates and microscale soil heterogeneity by dictating OM accessibility to microorganisms. However, traditional size and/or density fractionation techniques disrupt aggregate architecture, eliminating spatial information needed to fully understand intra-aggregate OM distribution. To quantify three-dimensional OM spatial distribution and automate segmentation in X-ray microcomputed tomography (µCT) imaging without human annotation bias, we developed an iodine gas vapor (I2) based staining workflow that eliminates labor-intensive manual annotation while maintaining segmentation accuracy, using aggregates from four taxonomically diverse soils (Xerofluvent, Haploxeroll Sphagnofibrist, Palehumult) with an 8-fold range of soil organic carbon. Human annotation of 10 µCT slices by the experienced and inexperienced annotators resulted in variations up to 3% in the Dice similarity coefficient (DSC), reflecting a degree of inherent subjectivity of manual labeling. Such inconsistencies are expected to compound as the number of manually annotated slices increases. Dual-energy µCT imaging at 33.1 keV (below the iodine (I) K-edge) and 33.2 keV (above the I K-edge) was used to resolve aggregate microstructure following I2 staining. The automated image subtraction pipeline identified OM regions by the I Kedge induced brightness increases, achieving DSC values of 0.58–0.83 relative to an experienced annotator. Sensitivity analyses revealed that the reconstruction alpha value—optimized via the open-source tool TomocuPy—and the 3D registration slice count were the primary determinants of accuracy, providing a novel benchmark for dual-energy soil imaging. The pipeline without GPU acceleration achieved 9.6 to 43.2 times faster than manual annotation. Using GPU-accelerated image post-processing and affine transformation matrices, the pipeline successfully segmented OM elements for large-scale datasets (3232×3232 pixel, 2048 slices) within ~5200 s from raw file acquisition to segmented output. The high-throughput approach enables the quantification of OM spatial distribution across diverse and heterogeneous soil.

Soil microbial biomass↗

PlasmidMaker: a Versatile, Automated, and High Throughput End-to-End Platform for Plasmid Construction

Plasmids are used extensively in basic and applied biology. However, design and construction of plasmids, specifically the ones carrying complex genetic information, remains one of the most time-consuming, labor-intensive, and rate-limiting steps in performing sophisticated biological experiments. Here, we report the development of a versatile, robust, automated end-to-end platform named PlasmidMaker that allows error-free construction of plasmids with virtually any sequences in a high-throughput manner. This platform consists of a most versatile DNA assembly method using Pyrococcus furiosus Argonaute ( Pf Ago)-based artificial restriction enzymes, a user-friendly frontend for plasmid design, and a backend that streamlines the workflow and integration with a robotic system. As a proof of concept, we used this platform to generate ~100 plasmids from six different species ranging from 5 to 18 kb in size from up to 11 DNA fragments within 3 days. PlasmidMaker should greatly expand the potential of synthetic biology.

Enghiad, Behnam↗

PlasmidMaker is a versatile, automated, and high throughput end-to-end platform for plasmid construction

Abstract Plasmids are used extensively in basic and applied biology. However, design and construction of plasmids, specifically the ones carrying complex genetic information, remains one of the most time-consuming, labor-intensive, and rate-limiting steps in performing sophisticated biological experiments. Here, we report the development of a versatile, robust, automated end-to-end platform named PlasmidMaker that allows error-free construction of plasmids with virtually any sequences in a high throughput manner. This platform consists of a most versatile DNA assembly method using Pyrococcus furiosus Argonaute ( Pf Ago)-based artificial restriction enzymes, a user-friendly frontend for plasmid design, and a backend that streamlines the workflow and integration with a robotic system. As a proof of concept, we used this platform to generate 101 plasmids from six different species ranging from 5 to 18 kb in size from up to 11 DNA fragments. PlasmidMaker should greatly expand the potential of synthetic biology.

59 BASIC BIOLOGICAL SCIENCES↗

Neural correspondence to spectrum of environmental uncertainty in multiple-cue probability judgment system with time delay

Despite state-of-the-art technologies like artificial intelligence, human judgment is critically essential in cooperative systems, such as the multi-agent system (MAS), which collect information among agents based on multiple-cue judgment. Human agents can prevent impaired situational awareness of automated agents by confirming situations under environmental uncertainty. System error caused by uncertainty can result in an unreliable system environment, and this environment affects the human agent, resulting in non-optimal decision-making in MAS. Thus, it is necessary to know how human behavior is changed to capture system reliability under uncertainty. Another issue affecting MAS is time delay, which can delay agent information transfer, resulting in low performance and instability. However, it is difficult to find studies on the influence of time delay on human agents. This study is about understanding the human decision-making process under a specific system reliability environment by uncertainty with time delay. We used concepts of expected and unexpected uncertainty to implement reliability of the system usage environment with three types of time delay conditions: no time delay, regular time delay, and irregular time delay conditions. We used electroencephalogram (EEG) for human cognitive neural mechanisms in multiple-cue judgment systems to understand human decision-making. In the reliability of system usage environment, the unreliable system environment significantly creates less memory load by less utilization of system rules for decision-making. In terms of time delay, delayed information delivery does not significantly affect memory load for decision-making.

cognitive process↗

Models, data, and scripts associated with “Prediction of Distributed River Sediment Respiration Rates using Community-Generated Data and Machine Learning”

This data package is associated with the publication “Prediction of Distributed River Sediment Respiration Rates using Community-Generated Data and Machine Learning’’ submitted to the Journal of Geophysical Research: Machine Learning and Computation (Scheibe et al. 2024). River sediment respiration observations are expensive and labor intensive to obtain and there is no physical model for predicting this quantity. The Worldwide Hydrobiogeochemisty Observation Network for Dynamic River Systems (WHONDRS) observational data set (Goldman et al.; 2020) is used to train machine learning (ML) models to predict respiration rates at unsampled sites. This repository archives training data, ML models, predictions, and model evaluation results for the purposes of reproducibility of the results in the associated manuscript and community reuse of the ML models trained in this project. One of the key challenges in this work was to find an optimum configuration for machine learning models to work with this feature-rich (i.e. 100+ possible input variables) data set. Here, we used a two-tiered approach to managing the analysis of this complex data set: 1) a stacked ensemble of ML models that can automatically optimize hyperparameters to accelerate the process of model selection and tuning and 2) feature permutation importance to iteratively select the most important features (i.e. inputs) to the ML models. The major elements of this ML workflow are modular, portable, open, and cloud-based, thus making this implementation a potential template for other applications. This data package is associated with the GitHub repository found at Please see the file level metadata (flmd; “sl-archive-whondrs_flmd.csv”) for a list of all files contained in this data package and descriptions for each. Please see the data dictionary (dd; “sl-archive-whondrs_dd.csv”) for a list of all column headers contained within comma separated value (csv) files in this data package and descriptions for each. The GitHub repository is organized into five top-level directories: (1) “input_data” holds the training data for the ML models; (2) “ml_models” holds machine learning models trained on the data in “input_data”; (3) “scripts” contains data preprocessing and postprocessing scripts and intermediate results specific to this data set that bookend the ML workflow; (4) “examples” contains the visualization of the results in this repository including plotting scripts for the manuscript (e.g., model evaluation, FPI results) and scripts for running predictions with the ML models (i.e., reusing the trained ML models); (5) “output_data” holds the overall results of the ML model on that branch. Each trained ML model resides on its own branch in the repository; this means that inputs and outputs can be different branch-to-branch. Furthermore, depending on the number of features used to train the ML models, the preprocessing and postprocessing scripts, and their intermediate results, can also be different branch-to-branch. The “main-*” branches are meant to be starting points (i.e. trunks) for each model branch (i.e. sprouts). Please see the Branch Navigation section in the top-level README.md in the GitHub repository for more details. There is also one hidden directory “.github/workflows”. This hidden directory contains information for how to run the ML workflow as an end-to-end automated GitHub Action but it is not needed for reusing the ML models archived here. Please the top-level README.md in the GitHub repository for more details on the automation.

13C↗