Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Information Automation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

Photogrammetry: Develop methods for detecting deviations from expected geometry of components in a glovebox environment (Lawrence Livermore National Laboratory Aging and Lifetimes Program FY23, Milestone 8651, Grading Criterion #4)

Measurements in glovebox environments are challenging to conduct. Currently deployed techniques are time consuming and provide limited information. These limitations make current disposition and future assessments of products challenging. In FY23 we developed methods for detecting deviations of components in a glovebox environment that address these shortcomings by rapidly collecting information-dense measurements. Conducted tests demonstrate that camera pose (position and orientation) can be derived for images utilized in the photogrammetry process. This information, in conjunction with common image processing techniques, enables automated detection and size estimation of surface features. First, image processing is used to identify marks, or localized regions, that standout from the surrounding area. Then, the physical size of the region can be estimated because scale can be determined for photogrammetry images. Development of this methods shows that photogrammetry can identify measurable deviations from expected geometries. Although measurement error when using this technique through a glovebox window still needs to be reduced, this capability shows promise for eliminating or reducing the use of tedious manual measurements.

42 ENGINEERING↗

Machine Learning Based Network Parameter Estimation Using AMI Data

The expansion of distribution power system and the growing penetration of distributed energy resources present new challenges for situational awareness. Calibrating the extended system model with sensor measurements and maintaining the usability is critical for utilities. This paper presents a distribution network parameter estimation (DNPE) approach using machine learning (ML) and metering data that improve the quality of extended distribution power system modeling. The reliability model can improve the ability of endpoint data to be translated into network-level situational awareness in real time and help distribution system operators (DSOs) solve branch flow and voltage problems. In addition, a data analytic and automate processing scheme is proposed to improve the sensor data quality and prevent misleading information. The effectiveness of the proposed method is verified with actual advanced metering infrastructure (AMI) data on a real utility feeder model, while considering the higher penetration of photovoltaic power generation. The test of DNPE and study results are demonstrated in this paper.

Parameter estimation, machine learning, power dist↗

The CanBikeCO Mini Pilot: Procedure and Preliminary Results

In fall 2020, the Colorado Energy Office, as part of the State of Colorado's "Can Do Colorado" initiative, initiated a project aimed at encouraging energy-efficient transportation during the COVID-19 pandemic. The initial mini-pilot provided e-bikes to 13 low-income households under an individual ownership model. This report assesses the impact of providing this additional mobility option on the travel behavior of participants. It also outlines the lessons learned from deploying a continuous monitoring platform to track the travel behavior. These lessons will influence the evaluation component for the full pilot, which will cover multiple geographic regions, start in summer 2021, and run for 2 years. The continuous data collection was enabled by a customized version of the open-source e-mission platform, called CanBikeCO, configured with a behavioral gamification feature. The Colorado Energy Office used this system to collect a unique data set consisting of 3 months of partially automated travel diaries, combining sensed and surveyed data and linked with demographic information, from 12 participants. The data collection process worked well overall: users generally liked the app, appreciated the game, and did not complain about battery life. The long tracking period introduced behavioral challenges in user engagement, which we plan to address using repeated patterns and automated status checks for the full pilot. The analysis results, based on the subset of trips with user-reported labels (68%), indicate that the e-bike was the dominant commute mode share (31%), in sharp contrast to the census bicycle commute mode share (<1%). E-bike trips primarily replaced single-occupancy vehicle (SOV) trips (28%), followed closely by walking (24%) and regular bike (20%). The non-motorized mode replacement corresponds to lower travel time and increased productivity enabled by the program. The emissions impact analysis of the program, computed using trip-level energy intensity factors, indicates savings of 1,367 lbs. of CO2. Although the results are strongly positive, the narrow demographic profile of study participants, their limited mobility alternatives, and nonuniform labeling indicate caution in broader interpretation. These preliminary results do suggest that such programs, supported by real-time education and support from program managers, can simultaneously meet equity and sustainability goals. The planned full pilot, addressing the data collection challenges and broadening the geographic scope, will provide additional insights into the generality of this approach.

ADVANCED PROPULSION SYSTEMS↗

Automating methods for estimating metabolite volatility

The volatility of metabolites can influence their biological roles and inform optimal methods for their detection. Yet, volatility information is not readily available for the large number of described metabolites, limiting the exploration of volatility as a fundamental trait of metabolites. Here, we adapted methods to estimate vapor pressure from the functional group composition of individual molecules (SIMPOL.1) to predict the gas-phase partitioning of compounds in different environments. We implemented these methods in a new open pipeline called volcalc that uses chemoinformatic tools to automate these volatility estimates for all metabolites in an extensive and continuously updated pathway database: the Kyoto Encyclopedia of Genes and Genomes (KEGG) that connects metabolites, organisms, and reactions. We first benchmark the automated pipeline against a manually curated data set and show that the same category of volatility (e.g., nonvolatile, low, moderate, high) is predicted for 93% of compounds. We then demonstrate how volcalc might be used to generate and test hypotheses about the role of volatility in biological systems and organisms. Specifically, we estimate that 3.4 and 26.6% of compounds in KEGG have high volatility depending on the environment (soil vs. clean atmosphere, respectively) and that a core set of volatiles is shared among all domains of life (30%) with the largest proportion of kingdom-specific volatiles identified in bacteria. With volcalc , we lay a foundation for uncovering the role of the volatilome using an approach that is easily integrated with other bioinformatic pipelines and can be continually refined to consider additional dimensions to volatility. The volcalc package is an accessible tool to help design and test hypotheses on volatile metabolites and their unique roles in biological systems.

59 BASIC BIOLOGICAL SCIENCES↗

Autonomous Multistate Nanoencoding Using Combinatorial Ferroelectric Closure Domains in BiFeO 3

Recent advances in ferroic materials have identified topological defects as promising candidates for enabling additional functionalities in future electronic systems. The generation of stable and customizable polar topologies is needed to achieve multistates that enable beyond-binary device architectures. Here, in this study, we show how to autonomously pattern on-demand highly tunable striped closure domains in pristine rhombohedral-phase BiFeO 3 thin films through precise scanning of a biased atomic force microscopy tip along carefully designed paths. By employing this strategy, we generate and manipulate closed-loop structures with high spatial resolution in an automated manner, allowing the creation of highly tunable and intricate topological domain structures that exhibit distinct polarization configurations without the need for electrode deposition or complex heterostructure growth. As a proof-of-concept for ferroelectric beyond-binary memory devices, we use such topological domains as multistates, engineering an alphabet and automating the symbolic writing/reading process using autonomous microscopy. The resulting information density is compared with that of current commercially available memory devices, demonstrating the potential of ferroelectric topological domains for multistate information storage applications.

BiFeO3↗

Generating Synthetic Time Series Photovoltaic Data with Real-World Physical Challenges and Noise for Use in Algorithm Test and Validation

The PV Fleet Data Initiative and other projects seek the develop algorithms for automated analysis of PV time series data for extraction of statistical information and other parameters of the data such as degradation rates, soiling loss information, tracker performance, clipping or curtailment, system availability and other valuable information. While there is a vast body of PV data available for application of said extraction algorithms it is difficult to validate these algorithms because the true parameters to be extracted are not known. There has been a wide use of synthetic data in the literature for algorithm validation but this synthetic data is typically very bounded by the problem or topic at hand. The PV Fleet Data Initiative project has demonstrated that real time series PV data almost always includes a host of data quality and physical problems that, in reality, any automated PV abstraction algorithm must handle appropriately. For this reason, this work describes the development of a complex synthetic PV times series data set that includes data quality and physical problems that have been experienced in real world PV data. The various quality and physical problems are documented in the synthetic data so that users can test the validity of various PV extraction algorithms as well as develop new algorithms to solve problems this data set can support.

14 SOLAR ENERGY↗

Raspberry Pi–powered temperature monitoring of growth chamber microclimates

While controlled environments are desirable for growing and measuring plants, growth chambers and greenhouses typically have microclimates that impact plant growth, development, and stress responses. Furthermore, opening and closing the doors of a controlled environment introduces variation in the environment, especially at temperature extremes, affecting both the measurements and the organisms within. Using multiple temperature data loggers to normalize results can be cost-prohibitive and rarely offers real-time feedback on temperature status. We used low-cost single-board computers, cameras, and temperature sensors to manage and capture growth chamber temperatures while acquiring plant image data. Detailed here are methods to document microclimates within a growth chamber so that data can be normalized to measured temperature information. This protocol describes a low-cost method for automated temperature monitoring, which enables both high-throughput measurements of temperature along with plant growth and stress responses via plant imaging.

Plant Sciences↗

Validating automated resonance evaluation with synthetic data

The integrity and precision of nuclear data are crucial for a broad spectrum of applications, from national security and nuclear reactor design to medical diagnostics, where the associated uncertainties can significantly impact outcomes. A substantial portion of uncertainty in nuclear data originates from the subjective biases in the evaluation process, a crucial phase in the nuclear data production pipeline. Recent advancements indicate that automation of certain routines can mitigate these biases, thereby standardizing the evaluation process and enhancing reproducibility. This research aims to provide a methodology, framework, and metrics for the validation of automated nuclear data evaluation software leveraging high-quality synthetic data that closely mimic real experimental observables. An introduced error metric provides a scale and intuitive measure of the evaluation quality by quantifying the estimate’s accuracy and performance across the specified energy range. Synthetic data provides access to experimental observables and underlying resonance parameters, enabling comparison of different evaluations. The methodology is demonstrated using Ta-181 isotope data in the resolved resonance region. The Automated Resonance Identification Subroutine (ARIS), which operates without prior resonance information, was used to test and showcase the framework’s capabilities utilizing the proposed error metrics. The results demonstrate the effectiveness of the proposed approach and framework for optimizing software parameters and testing hypotheses through “what-if” controlled experiments, such as modifying assumptions about experimental conditions or average resonance parameters.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Enabling site-specific well leakage risk estimation during geologic carbon sequestration using a modular deep-learning-based wellbore leakage model

Geologic carbon sequestration (GCS) is a promising technology for mitigating net carbon emissions and growing climate concern by storing CO 2 in reservoirs. Oil and gas brownfields are an attractive option for CO 2 storage, but these sites have many historical wellbores from petroleum production and can be a potential leakage pathway for CO 2 or formation brine. Therefore, risk management of GCS operations requires an assessment of potential well leakage. Due to the high uncertainty of the system, stochastic approaches are ideal for quantifying the range of risk behaviors, but they must be computationally efficient in the face of complex physics. Here, we develop a new physics-centric deep learning wellbore model to predict the leakage of CO 2 and brine through leaky wellbores. Multi-physics numerical simulations were used to generate data sets, and physics-informed features were introduced. Neural networks were optimized with an automated searching algorithm. Feature analysis quantifies the impact of each feature on model prediction and confirms the role of physics-inspired parameters. The model shows high predictive performance across a wide range of geologic and injection conditions and well attributes. In conclusion, a case study illustrates how the model is applied to assess well leakage in GCS operations.

58 GEOSCIENCES↗

Uncertainty-Aware Machine Learning for Small-Angle X-ray Scattering Analysis in Autonomous Experimentation

Small-angle X-ray scattering (SAXS) is a powerful high-throughput characterization tool for probing nanoscale structure in native sample environments, providing real-time morphological information such as nanoparticle size and shape during synthesis. However, automated SAXS data analysis for extracting meaningful structural parameters is non-trivial and remains a bottleneck in closed-loop experimentation towards autonomous materials discovery, which demands fast, reliable, and uncertainty-aware data analysis. Here, we develop a machine-learning approach for automated SAXS analysis tailored to closed-loop nanoparticle synthesis. A Random Forest (RF) regression model is trained on 100,000 synthetic SAXS curves generated from polydisperse spherical nanoparticles with realistic background contributions. Using normalized one-dimensional SAXS intensity profiles as input, the RF model directly predicts nanoparticle radius, size polydispersity, and background parameters, while the ensemble standard deviation across trees provides built-in uncertainty quantification (UQ). On synthetic data, we show that combining fit-quality metrics (R 2 , MAE) with thresholds on prediction uncertainty reliably identifies accurate parameter estimates without access to ground truth. We then apply the trained model to 365 experimental SAXS profiles of citrate-reduced gold nanoparticles synthesized using an automated droplet-flow microreactor with in situ SAXS at a synchrotron beamline, classifying the results into high- and low-confidence subsets based on UQ metrics. Finally, we integrate RF-based SAXS analysis into a simulated closed-loop optimization campaign using Gaussian process Bayesian optimization to minimize nanoparticle polydispersity, benchmarking against conventional automated Levenberg–Marquardt fitting. The RF-guided campaign exhibits substantially faster convergence and lower relative opportunity cost (∼0.07 vs ∼0.3), demonstrating that uncertainty-aware machine-learning SAXS analysis significantly enhances the efficiency and robustness of autonomous nanomaterials synthesis workflows.

Bayesian optimization↗

Concept of Operations of Next-Generation Traffic Control Utilizing Infrastructure-Based Cooperative Perception

This paper provides a system architecture for an infrastructure-based cooperative perception fusion engine for next-generation traffic control. This engine will provide a complete state-space digital representation with measurable accuracy to support a wide-range of applications. The architecture includes inputs, functional flow, data standardization recommendations, outputs, and supported applications. The cooperative perception engine addresses critical needs with respect to accelerating the benefits of automation through intelligent roadway infrastructure, which complements and accelerates connected and automated vehicle (CAV) technology. The cooperative perception acquires and fuses information from sensors (radar, LiDAR, and cameras) and CAVs to perceive roadway traffic states of moving objects, creates a complete 3D digital representation of that state-space, and communicates it to downstream application such as intelligent signal control, safety and energy applications, and cooperate driving applications. The intelligent roadway infrastructure approach, as opposed to a vehicle-centric approach, is more scalable because it can be deployed to the roughly 300,000 signalized intersections more readily than over 300 million vehicles in the United States, and accrues early-stage benefits equitable to all roadway users addressing safety, equity, fuel efficiency, and greenhouse gas reduction.

ADVANCED PROPULSION SYSTEMS↗

Concept of Operations of Next-Generation Traffic Control Utilizing Infrastructure-Based Cooperative Perception: Preprint

This paper puts forth a system architecture for an infrastructure-based cooperative perception (CP) fusion engine, to provide a complete state-space digital representation, with measurable accuracy, to support a wide-range of applications. The architecture includes the inputs, functional flow, data standardization recommendations, outputs and supported applications. The CP engine addresses critical needs with respect to accelerating the benefits of automation through intelligent roadway infrastructure (IRI), that complements and accelerates connected and automated vehicle (CAV) technology. that the CP acquires and fuses information from sensors (radar, LiDAR, and cameras), and CAVs to intelligently perceive roadway traffic states of all moving objects, create a complete three-dimensional digital representation of that state-space, and communicate it to downstream application such as intelligent signal control, safety and energy applications, and cooperate driving applications for CAVs as examples. The IRI approach, as opposed to a vehicle centric approach, is found to be more scalable in that it can deployed to the roughly 300,000 signalized intersections more readily than the over 300 million vehicles in the US, and accrues early-stage benefits equitable to all roadway users addressing safety, equity, fuel efficiency, and GHG reduction.

ADVANCED PROPULSION SYSTEMS↗

AI for Interpreting Nuclear Power Plant Documents for Power Uprates

To reduce the cost and time needed for regulatory compliance, nuclear power plants (NPPs) can utilize artificial intelligence (AI) to assist in interpreting complex and voluminous documents that typically span thousands of pages. Usually, the process of interpreting a plant’s technical specifications (TSs) and associated documents is labor intensive. This study aims to understand what processes state-of-the-art large language models (LLMs) can automate and to identify the pitfalls associated with using LLMs to reduce human labor costs and time. This research uses a recent AI technology called retrieval augmented generation (RAG), which retrieves pages of information from TSs and associated documents to assist with NPP power uprates (cleared to produce more power). LLMs are integral to RAG because they create human-like responses based on the retrieved information, aiding in the interpretation and application processes. A baseline case demonstrates how LLMs can operate successfully for a power uprate application. Then five use cases show five types of potential failures: (1) RAG retrieving the incorrect information, (2) RAG misinterpreting the retrieved information, (3) RAG relying on knowledge not contained in the retrieved information, (4) RAG hallucinating, and (5) RAG refusing to answer. The results of the five use cases suggest that automating the human interpretation of TSs and associated documents with AI should be approached with caution. A subject-matter expert reviewed the AI outputs from the five use cases and concluded that an LLM can produce technical information that is needed to produce power uprate applications in certain instances.

21 - SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLAN↗

Automated detection of photovoltaic cleaning events: A performance comparison of techniques as applied to a broad set of labeled photovoltaic data sets

Extracting accurate soiling loss information from photovoltaic (PV) production data first requires segmenting the time series data per natural or manually occurring cleaning events. Maintenance logs are often incomplete, rain data are often unavailable, and the debate on rain thresholds for cleaning and dew or wind cleanings is still ongoing. The present work aims to overtake these issues by improving automated methods to detect these cleaning events and therefore improve extraction of soiling loss information. Time series power production data from 22 PV inverters were labeled for natural or manually occurring cleaning events. The data sets were carefully selected to include varying degrees of soiling, cleaning events, and noise. Several algorithms, including filtering logic and change point detection, were examined for efficacy at detecting the labeled cleanings. All the methods introduced except for changepoint detection showed significant improvement at detecting the labeled cleaning events per the mean F 1 score. Furthermore, the highest performing cleaning detection algorithm achieved an absolute increase in the mean F 1 score of 43% over the default version of the RdTools stochastic rate and recovery (SRR) algorithm. The highest performing algorithm included irradiance filtering and a cleaning detection threshold, adjusted based on the 40-day centered rolling median of the absolute day-to-day deviations in the daily performance index (PI). Furthermore, these improvements are promising as cleaning detection is an essential step in the automated analysis of PV soiling.

14 SOLAR ENERGY↗

Real‐time XFEL data analysis at SLAC and NERSC: A trial run of nascent exascale experimental data analysis

X‐ray scattering experiments using free electron lasers (XFELs) are a powerful tool to determine the molecular structure and function of unknown samples (such as COVID‐19 viral proteins). XFEL experiments are a challenge to computing in two ways: (i) due to the high cost of running XFELs, a fast turnaround time from data acquisition to data analysis is essential to make informed decisions on experimental protocols; (ii) data‐collection rates are growing exponentially, requiring new scalable algorithms. Here we report our experiences analyzing data from two experiments at the Linac Coherent Light Source (LCLS) during September 2020. Raw data were analyzed on NERSC's Cori XC40 system, using the Superfacility paradigm: our workflow automatically moves raw data between LCLS and NERSC, where it is analyzed using the software package CCTBX. We achieved real time data analysis with a turnaround time from data acquisition to full molecular reconstruction in as little as 10 min—sufficient time for the experiment's operators to make informed decisions. By hosting the data analysis on Cori, and by automating LCLS‐NERSC interoperability, we achieved a data analysis rate which matches the data acquisition rate. Completing data analysis within 10 min is a first for XFEL experiments and an important milestone if we are to keep up with data‐collection trends.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Towards non-contact pollution monitoring in sewers with hyperspectral imaging

Monitoring water quality in sewers is challenging, particularly because state-of-the-art technologies require contact with the raw wastewater. The presence of fat, oil, grease, and solids makes automated grab sampling difficult and causes sensor fouling. To overcome these limitations, non-contact methods based on light reflectance, such as hyperspectral imaging (HSI), are gaining attention. However, HSI has never been tested for raw wastewater. To assess its accuracy for measuring pollution, we developed a laboratory setup and performed targeted experiments with a combination of raw and diluted wastewater, as well as synthetic turbidity stock solutions. We measured seven pollution variables: chemical oxygen demand, turbidity, dissolved organic compounds, ammonium, total nitrogen, phosphate, and sulphates. We used automated pixel selection and partial least squares regression to retrieve pollution information from the hyperspectral images. Our results, based on 144 samples, suggest that HSI can estimate pollution levels with a precision in the range of state-of-the-art absorbance spectrophotometric methods. Additionally, we found that the combination of pixel and wavelength selection, enabled by the hyperspectral data structure, significantly influences the performance of partial least square modelling. Overall, our findings indicate that HSI is a promising technology for non-contact monitoring of water quality in raw wastewater.

54 ENVIRONMENTAL SCIENCES↗

Red supergiant candidates for multimessenger monitoring of the next Galactic supernova

ABSTRACT We compile a catalogue of 578 highly probable and 62 likely red supergiants (RSGs) of the Milky Way, which represents the largest list of Galactic RSG candidates designed for continuous follow-up efforts to date. We match distances measured by Gaia DR3, 2MASS photometry, and a 3D Galactic dust map to obtain luminous bright late-type stars. Determining the stars’ bolometric luminosities and effective temperatures, we compare to Geneva stellar evolution tracks to determine likely RSG candidates, and quantify contamination using a catalogue of Galactic AGB in the same luminosity-temperature space. We add details for common or interesting characteristics of RSG, such as multistar system membership, variability, and classification as a runaway. As potential future core-collapse supernova progenitors, we study the ability of the catalogue to inform the Supernova Early Warning System (SNEWS) coincidence network made to automate pointing, and show that for 3D position estimates made possible by neutrinos, the number of progenitor candidates can be significantly reduced, improving our ability to observe the progenitor pre-explosion and the early phases of core-collapse supernovae.

Astronomy & Astrophysics↗