Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Training Analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 217 records · Page 12

Verification and Validation of START: A Spent Nuclear Fuel Routing and Decision Support Tool

The Stakeholder Tool for Assessing Radioactive Transportation (START) is a web-based geospatial decision-support tool being developed by the US Department of Energy’s Office of Integrated Waste Management (IWM) to support federal interim storage for spent nuclear fuel (SNF) and associated transportation. START provides many functions for the IWM program including: serving as a communications tool for conveying geospatial data and information, an options analysis tool for exploring potential transport modes and routes for transporting SNF from nuclear power plants to future federal interim storage facilities, an emergency response planning tool for Tribes and States to identify training needs along potential SNF transport corridors, an environmental analysis tool for estimating potential radiation dose exposure from incident-free and incident-case SNF transport conditions, and a systems analysis support tool providing route-related inputs for system throughput analysis. As part of the START development process, a verification and validation (V&V) effort is being undertaken. In the initial V&V phase, several outputs of the START tool were checked such as the total distance, population and population densities within the buffer zone, and incident free dose. The V&V process is fluid as it will be utilized after each version change to ensure that the core functionalities of the tool are maintained and the results are consistent with the previous versions. Efforts have also been put into developing scripts to aid in the process of automating certain sections of the V&V work. As some of the V&V efforts use Environmental Systems Research Institute’s (ESRI) tools upon which the START framework is built, the START data were compared against outputs from tools like Quantum Geographic Information System (QGIS) for buffer zone populations and route lengths to ensure independence of the V&V process. Good agreement was observed between the START results and the independent V&V studies with the majority of the differences falling between 1% and 5% for populations within the buffer zone and route distance. This presentation describes the development and design of the START tool, V&V methods employed for various metrics of interest and their respective results, and future plans.

START, transportation, GIS, V&V↗

Summary of NDC Capacity Building Workshop and Regional Seismic Travel Time (RSTT) in combination with Data Sharing and Integration Training

The NNSA Seismic Cooperation Program (SCP) sponsored Stephen Myers (LLNL), Michael Begnaud (LANL), Brian Young (SNL) and Istvan Bondar (Research Center for Astronomy and Earth Sciences, Hungary) to serve as a presenters/trainers at the “NDC Capacity Building Workshop and Regional Seismic Travel Time (RSTT) in combination with Data Sharing and Integration Training” September 4-8 2022 in Muscat, Oman (See Appendix A for the agenda). The workshop and training (workshop from here forward) was organized by the Comprehensive Nuclear-Test-Ban Treaty Organization (CTBTO) Provisional Technical Secretariat (PTS). The first half of the week was devoted to NDC workshop activities, and the second half was devoted to RSTT training. Fifty-five participants from 27 countries and the CTBTO-PTS attended the 5-day workshop (See Appendix B for list of participants and countries of origin). Presentations from the PTS described the International Monitoring System (IMS), International Data Centre (IDC) products, and metrics of regional data utilization. Contributed presentations from each country’s scientists included descriptions of regional and national networks, methods of data analysis, and needs for material and technical assistance. Training included an overview of the RSTT method and instruction on how to locate seismic events with the iLoc program, which utilizes RSTT travel times to reduce bias in event location estimates. Methods of seismic tomography and the need for a high-quality tomographic set, including seismological “ground truth”, were emphasized. Seismological “ground truth” or “GT” is a term that has come to mean both events with known location and events with well-characterized locations that are estimated using seismological data, typically with epicenter accuracy of 5 km or better. Notably, the instructional platform has migrated from UNIX shell scripts to Jupyter Notebooks. Jupyter Notebooks have the advantage being more visually intuitive, including display of graphics within the notebook. Each notebook includes every processing step that participants need to reproduce the entire exercise.

58 GEOSCIENCES↗

Video Summarization Using Deep Action Recognition Features and Robust Principal Components Analysis

In an instance where desired pre-defined actions, behaviors, or other categories are known a priori, various video classification and recognition models can be trained to discover those classifications and their location within the video. Absent that information, one might still be tasked with identifying interesting portions within a video, a process which—if done manually—is onerous and time-consuming as it requires manual inspection of the video itself. Recognizing high-level interesting segments within a whole video has been a general area of interest due to the ubiquity of video data. However the size of the data makes storage, retrieval, and inspection of large collections of videos cumbersome. This problem motivates the task of generating shortened clips highlighting the primary content of a video, relieving the burden of having to watch the entire video. This paper presents an unsupervised method of creating shortened clips of videos, enabling the rapid review of the most interesting content within a video. Our method uses features extracted from pre-trained action recognition models as input to online moving window robust principal component analysis to generate summaries. The procedure is tested on a publicly available video summarization dataset and demonstrates comparable performance to state-of-the-art in an un-augmented setting while requiring no training.

Claborne, Daniel M.↗

MPEX AI Digital Twins

All magnetically confined plasma fusion power plant concepts (Tokamak, Spherical Tokamak, Stellarator, Mirror, ...) must exhaust the heat and plasma from the core confinement region to the material walls. The primary channel for this exhaust is through a plasma divertor which directs plasma along open magnetic field lines to a material target. The Material Plasma Exposure eXperiment (MPEX) illustrated in Figure 1, is a high-power, steady-state linear plasma device designed to produce the plasma material interaction (PMI) conditions of the divertor of future magnetic confinement fusion power plants: energy flux 20MW/m 2 , ion fluence 1031/m 2 , pulse duration 106 sec. These goals of plasma exposure in MPEX are well beyond those achieved in magnetic fusion experimental devices. Successfully achieving these high power steady state conditions for long pulses requires operational control of the heating and particle sources and the plasma flux to the walls and target. The MPEX AI Hot Spot Controller, proposed in this project, will help achieve the operational milestones of MPEX. The MPEX device will begin commissioning at the end of FY26. A smaller proto-MPEX was operated for 14,666 plasma discharges and will resume operation in September of 2025 as proto-MPEX-lite, with reduced capability, to test a new window for the Helicon plasma source. The proto-MPEX data has undergone surrogate modeling with machine learning methods (R. Archibald, 2022 IEEE International Conference on Big Data). This proto-MPEX data will be used to begin development of the AI digital twins described in this white paper. The scientific mission of MPEX is to qualify materials of different composition for use in the high energy and plasma flux conditions of a fusion power plant. The materials exposed in MPEX will in some cases be exposed to high neutron fluxes at other ORNL facilities to measure the changes to their PMI properties. The targets exposed in MPEX will be transported under vacuum to a Surface Analysis Station (SAS). The SAS will be equipped with the following diagnostics: Focused Ion Beam (FIB) for trench milling, 100-400 angstrom resolution scanning electron microscope (SEM), surface mapping x-ray spectrometer, high resolution camera, and a future upgrade to a laser induced breakdown spectroscopy quadruple mass spectrometer (LIBS-QMS). The MPEX experiments will generate diverse pre- and post-exposure measurement data of detailed material properties down to the crystal grain level in 3D for post-exposure assessment of PMI damage (e.g. cracking, melting, erosion and redeposition of the material). Physics models for the PMI, and how the material composition and manufacturing impact its performance under high energy plasma exposure, need to be validated with MPEX data to guide the selection of new candidate materials. Our vision for the MPEX AI Digital Twins project is to supply experimental and physics model simulation data to train Artificial Intelligence (AI) models for data processing, analysis, operational control, PMI and materials simulation to maximize the scientific output of the MPEX device. Ultimately, an AI digital twin of MPEX material assessment metrics for tested and synthetic material types with simulated PMI will be trained by the AI Modeling Teams on the experimental and physics simulation data submitted to the American Science Cloud by this project. A purely empirical search for the best material is inefficient given the finite number of samples that can be tested on MPEX. In order to expand the material properties database for training the MPEX Material Assessment AI Digital Twin, and to gain physics understanding of the PMI processes, physics models of the material properties and PMI processes are required. The physics simulations provide detailed simulation data, like impact angles for plasma ions, sputtering yields, transport of the ionized sputtered target material in the plasma, and redeposition locations. This simulation data expands the measurement data for deeper physics understanding. The experimental data is essential to validate the PMI and material structure simulation models. The validated models can then be used to generate new simulation data of MPEX material assessments for synthetic material compositions that have not been exposed in MPEX. These predictive simulations, plus the whole experimental dataset, will be used to train the MPEX Material Assessment AI Digital Twin allowing a rapid generative AI search for new materials with reduced PMI damage by interpolating the domain of the training set. These new optimum materials can be simulated with the physics codes and/or tested in MPEX. The ability of AI neural networks to interpolate multi-dimensional parameter spaces and generate virtual data is exploited for a more efficient search for optimum materials. The advent of the Transformational AI Models Consortium (TAIMC) is an opportunity to engage with state of the art private and public AI developers to achieve the goals of the AI digital twins and AI accelerated physics models proposed in this project. Our partners at ORNL from the Advance Scientific Computing Research (ASCR) organization will collaborate in accelerating the integrated plasma material interaction simulation framework. This simulation framework will provide a platform for generating simulation data across a range of physical fidelities, including hybrid methods that produce multi-fidelity results. This data will be leveraged for AI model development, both for generation of surrogates and the automation of simulation campaigns. A part of the research below will include collaborative efforts with the TAIMC to (i) adapt data storage approaches to ensure AI-readiness, (ii) provide a protypical exemplar to inform and exercise constructed workflows, and (iii) generate and share data, using the TAIMC unified AI data standard, for foundational models that will be trained from multiple sources across the DOE complex. We will also collaborate with the TAIMC, as well as the planned AI modeling teams, to develop approaches for reducing the cost of data generation. These include tailored multi-fidelity approaches as well as fine-tuning strategies to augment general, large-scale foundational models.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

MindSynchro

This report presents the developments and results of MindSynchro project as part of DOE OE FOA 1861. DOE and Pacific Northwest National Laboratory (PNNL) have made available to FOA awardees datasets containing years of real historical data recorded from various phasor measurement units (PMUs) which are installed in three large US interconnections: Texas (IC A), Western (IC B), and Eastern (IC C). The main goal of the project, which was successfully achieved, was to develop methods for detection and identification of events which are relevant for power grid operation. Tasks performed for achieving the project goals included data exploration and pre-processing, the development and application of physics-based features, data analysis and labeling based on unsupervised learning approaches, training and testing of DSSL models for classification of events which are relevant for power grid operation, and deployment of solutions to cloud environments. The methods developed in the project can potentially provide relevant benefits to power grid asset owners/operators in general in terms of situational awareness. Two main types of outcomes can be provided by these tools: Identification of specific relevant power grid event types: Semi-supervised ML methods developed in the project can adequately employ not only the relatively scarce labeled data but also the large amount of available unlabeled data to train models for detection of specific event types. Such methods enable the application of trained models for the detection of events in a population of PMUs much larger than that associated to the labeled events. Support in data labeling / label validation: Labels are critical for training of models for identification of specific types of events. However, labeling large amounts of data is a manual and tedious process. This means that such process is error prone and is not scalable. Methods developed in the project, based on ensembles of clustering models, have been successfully employed for turning manual labeling into a scalable process. Accurate identification of specific relevant events can provide the operators with immediate situational awareness that could otherwise require hours or days of analysis from domain experts. We envision that such methods could be initially employed in support of post-mortem analysis of events and, as confidence is gained, they could be employed for online/real-time support, providing, among other benefits, insights for avoiding major events which could happen due to a combination of smaller ones. On the longer term, related methods could potentially be employed to improve protection and control.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Process Considerations for the Production of Hydrogen via Steam Reforming of Oxygenated Gases from Biomass Pyrolysis and Other Conversion Processes

Abstract In 2021, average CO 2 emissions was 9.7 g CO 2 /g H 2 produced, primarily based on steam methane reforming (SMR) technology that currently dominates hydrogen production. The substitution of natural gas (NG) and other fossil feedstocks in SMRs with renewable gases may be considered as an option for reducing greenhouse gas emissions. This analysis explores process impacts and constraints associated with the potential introduction of biogenic gases into SMRs. The results indicate that replacing NG in the fuel train of the SMR, followed by partial replacement of NG in the feed train may be a feasible approach. For the CO‐ and CO 2 ‐rich gas compositions assumed in this analysis the results indicate that feed train may accommodate up to 25 mole % of biogenic gases using allowances in existing designs and/or with small modifications, while maintaining similar hydrogen output. NG substitution in higher proportions require more major changes because of increased flow rates and heat exchange requirements in the system. Biogenic gases with lower CO 2 and higher calorific values are advantageous for NG substitution, and dry reforming using the CO 2 present in the feed gas can reduce steam consumption and increase process efficiency within limits where coking does not become a new constraint.

08 HYDROGEN↗

Understanding How Organizations Handle Cybersecurity

If there is anything we can learn from the media, it is the frequency and severity of cyber-attacks is increasing and there are not enough qualified people to combat the risk organizations are facing. Current estimates say there are 3.5 million available cybersecurity related jobs globally and there has been a 350% growth in cybersecurity jobs since 2013 (Group, 2020). The Idaho Cyber Research Project (ICRP) is focused on finding an implementing solution to the problems in the workforce development pipeline. Our team consists of Cohort 2 of the ICRP, we are tasked with solving issues faced by organizations hiring new cyber personnel. To provide solutions to these issues we focused our research on four components of workforce availability and competency: resume and transcript analysis, apprenticeships, cyber incident response plan development, and adversarial mindset training. From this research we have produced the following focus areas and subsequent steps for each component of workforce capability: transcript and knowledge skills abilities (KSA) focused analysis, cybersecurity apprenticeships programs, the value of an adversarial mindset, and a guide to setting up cyber incident response plans for underprepared organizations. These solutions can be further developed and implemented to reduce the gap in workforce demand and talent.

97 MATHEMATICS AND COMPUTING↗

Search for HH → bbτ⁺τ⁻ Using Run 3 Scouting Data Analyze b-tagging and tau-tagging Performance with Unified Particle Transformer

B-tagging and tau-tagging performances play an important role in the search for the rare event HH → bbτ⁺τ⁻. A transformer-based neural network, Unified Particle Transformer, is applied for both tagging tasks, and Run 3 proton–proton collision scouting data at center-of-mass energy of 13.6 TeV is used. The scouting data stream accepts events at a much higher rate compared to traditional triggers, but stores only the objects reconstructed in the trigger, no low-level detector information. Therefore, existing taggers trained for the offline event reconstruction cannot be used. Analysis of the SoftMax plots, ROC/AUC curves, confusion matrix, accuracy and losses are used to evaluate model performance. Specifically, the tagging efficiency of the signal and misidentification probability across multiple background processes are compared for varying working points. Different training samples with distinct distributions of jet flavors are utilized and related model performances are analyzed. Interpretability methods, such as Integrated Gradients, may further be applied to study the input features’ influence on the model’s decisions, providing insights into potential improvements.

Chen, Blair [Purdue U., West Lafayette; Fermilab]↗

Trajectory Optimization via Unsupervised Probabilistic Learning On Manifolds

This report investigates the use of unsupervised probabilistic learning techniques for the analysis of hypersonic trajectories. The algorithm first extracts the intrinsic structure in the data via a diffusion map approach. Using the diffusion coordinates on the graph of training samples, the probabilistic framework augments the original data with samples that are statistically consistent with the original set. The augmented samples are then used to construct conditional statistics that are ultimately assembled in a path-planing algorithm. In this framework the controls are determined stage by stage during the flight to adapt to changing mission objectives in real-time. A 3DOF model was employed to generate optimal hypersonic trajectories that comprise the training datasets. The diffusion map algorithm identfied that data resides on manifolds of much lower dimensionality compared to the high-dimensional state space that describes each trajectory. In addition to the path-planing worflow we also propose an algorithm that utilizes the diffusion map coordinates along the manifold to label and possibly remove outlier samples from the training data. This algorithm can be used to both identify edge cases for further analysis as well as to remove them from the training set to create a more robust set of samples to be used for the path-planing process.

42 ENGINEERING↗

Weakly supervised anomaly detection for resonant new physics in the dijet final state using proton-proton collisions at $\sqrt{s}$ = 13 TeV with the ATLAS detector

An anomaly detection search for narrow-width resonances beyond the Standard Model that decay into a pair of jets is presented. The search is based on 139 fb −1 of proton-proton collisions at $\sqrt{s}$ = 13 TeV recorded during 2015–2018 with the ATLAS detector at the Large Hadron Collider. The analysis is optimized without a particular signal model and aims to be sensitive to a broad range of new physics. It uses two different machine learning strategies to estimate the background in different signal regions. In each region, a weakly supervised classifier is trained to distinguish this background model from data. The analysis focuses on events with high transverse momentum jets reconstructed as large-radius jets. The mass and substructure of these jets are used as inputs to the classifiers. After a classifier-based selection, the distribution of the invariant mass of the two jets is used to search for potential local excesses. The model-independent results of both the anomaly detection methods show no signs of significant local excesses. In addition to model-independent results, a representative set of signal models is injected into the data, and the sensitivity of the methods to these scenarios is reported.

Aad, G. [Aix-Marseille Université] (ORCID:00000002↗

Nuclear Criticality Safety Training: Needs and Efforts [Slides]

This presentation on Nuclear Criticality Safety Training: Needs and Efforts covers several points. This presentation starts with a quick overview of Nuclear Criticality Safety (NCS) Training. Then it lists current and ongoing training with a focus on the training in progress. This presentation then concludes with a look at supporting efforts.

96 KNOWLEDGE MANAGEMENT AND PRESERVATION↗

Toward designing effective exascale scientific computing workflows: experiences and best practices

Many fields within scientific computing have embraced advances in big-data analysis and machine learning, which often requires the deployment of large, distributed and complicated workflows that may combine training neural networks, performing simulations, running inference, and performing database queries and data analysis in asynchronous, parallel and pipelined execution frameworks. Such a shift has brought into focus the need for scalable, efficient workflow management solutions with reproducibility, error and provenance handling, traceability, and checkpoint-restart capabilities, among other needs. Here, we discuss challenges and best-practices for deploying exascale-generation computational science workflows on resources at the Oak Ridge Leadership Computing Facility (OLCF). We present our experiences with large-scale deployment of distributed workflows on the Summit supercomputer, including for bioinformatics and computational biophysics, materials science, and deep learning model optimization. We also present problems and solutions created by working within a Python-centric software base on traditional HPC systems, and discuss steps that will be required before the convergence of HPC, AI, and data science can be fully realized. Our results point to a wealth of exciting new possibilities for harnessing this convergence to tackle new scientific challenges.

Coletti, Mark↗

EXERGETIC: De-Risking Next-Generation Resilient Geothermal Hybrids via At-Scale Evaluation Using Virtual Emulation Digital Twin Environment for Efficient Operation

The DOE-GTO-funded project, award number 5.1.2.12, entitled "EXERGETIC - De-risking Next Generation Resilient Geothermal Hybrids via at-Scale Evaluation Using a Virtual Emulation Digital Twin Environment for Efficient Operation," advances the solution to these challenges by developing and validating a geothermal co-emulation environment implemented at the National Laboratory of the Rockies (NLR)'s Advanced Research on Integrated Energy Systems (ARIES) platform. This framework enables the de-risking of next-generation geothermal and geothermal hybrid systems through high-fidelity modeling, real-time digital emulation, advanced control strategies, and techno-economic assessment. The project focused on geothermal hybrid configurations that integrate geothermal power plants with concentrated solar power and underground thermal energy storage, enabling enhanced efficiency, flexibility, and grid support capabilities. The main goal of this project was the development of a geothermal digital co-emulation environment to demonstrate the technical and economic value of geothermal hybrid systems and their contribution to grid stability and flexibility. The EXERGETIC framework combined physics-based models, controls, and real assets at ARIES, including digital real-time simulators (DRTS), a 20-MW-scale controllable grid interface (CGI), and a 2-MW conventional generator. Detailed transient models were developed for the key subsystems of a hybrid geothermal plant, including parabolic trough solar collectors, reservoir thermal energy storage (RTES), and a binary Organic Rankine Cycle (ORC) power plant. The ORC model explicitly captured thermal inertia and off-design operation and integrated control strategies to dynamically respond to electric load profiles. The models were validated against published experimental and numerical studies, demonstrating strong agreement and confirming the accuracy and robustness of the modeling approach. The resulting digital twin represents geothermal-solar-storage systems at multiple scales (1 MW to 100 MW) and enables realistic emulation of grid-connected operation. The control architecture allows the geothermal resource to provide stable baseload generation, while solar and stored thermal energy supply flexible, dispatchable support during periods of high demand or variable grid conditions. A key contribution of the EXERGETIC project is the demonstration that geothermal hybrid systems can be designed to be active grid assets rather than passive baseload generators. Using the ARIES platform, the digital twin was evaluated under multiple grid scenarios, including load following, voltage support at the distribution level, and frequency response at the transmission level. Results show that hybrid geothermal systems can respond effectively to dynamic grid conditions, providing inertia-like behavior, primary frequency support, and voltage regulation through coordinated control. In addition to the performance and grid services capability analysis of geothermal and hybrid geothermal systems, the EXERGETIC project also focused on scalability and techno-economic analysis of geothermal hybrid plants. In particular, for the scalability analysis, machine-learning (ML)-based surrogate models were trained using data generated from the geothermal digital twin under different grid-connected scenarios and plant capacities. These ML models demonstrated strong interpolation and extrapolation capabilities across plant sizes, accurately reproducing both steady-state and transient responses with very low errors. Regarding the techno-economic analysis, plant performance results were integrated with cost models for hybrid geothermal systems, and the levelized cost of electricity (LCOE) was used as the main economic metric to evaluate system performance across a range of system capacities, solar shares, solar multiples, and storage durations. Results indicate that economies of scale significantly reduce geothermal LCOE as plant capacity increases, with large-scale systems (25-100 MW) achieving substantially lower costs than small plants. Hybridization with solar thermal energy and storage further improves economic performance by increasing capacity utilization and enabling flexible dispatch. In addition, thermal storage plays a critical role in reducing LCOE by maximizing geothermal, solar, and stored energy resources. In summary, the results from this project demonstrate that geothermal hybrid systems represent a promising alternative for increasing the energy conversion efficiency of geothermal technologies, contributing to the preservation of geothermal resources, and supporting the transition of geothermal plants from traditional baseload resources into flexible, resilient, and cost-competitive energy conversion technologies.

15 GEOTHERMAL ENERGY↗

Improving Fission Products at CARIBU: Near Field Detection (Q3/FY21 Quarterly Progress Report)

The overview of our project as well as most recent results have been presented at the NSARD review in April, and more recently during the WoNDRAM workshop. We have been dealing with recent personnel changes in the group. Miguel Bencomo, who has worked with us on this project as a postdoc at LLNL since 2019, will be terminating his appointment within the next two weeks to take a full-time appointment with Raytheon. We continue training our new team member Dan Hoff to take over data analysis from the last experiment we performed, as well as prepare for upcoming measurements. Additionally, we are working with a summer student, John Wilkinson, who is training to perform GEANT4 simulations of beta detector this summer.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

When physics-informed data analytics outperforms black-box machine learning: A case study in thickness control for additive manufacturing

Aerosol jet printing (AJP) has emerged as a promising noncontact additive manufacturing method for high-resolution printing for a wide range of material systems. A key challenge limiting the broader adoption of AJP in the material science community is the lack of methods to precisely control thickness. Herein, we develop a model-based design of experiment (MBDoE) framework that integrates physics-informed models, nonlinear regression, and information criteria to postulate, select and calibrate the best model to describe and optimize the AJP manufacturing process. Starting with already available data from system commissioning (e.g., prior single variable sensitivity analysis), four candidate physics-informed models are postulated and trained. MBDoE identifies a single additional optimal experiment to validate these predictive models with quantified uncertainties, which are then used to determine the best experimental conditions to control printed film thickness. As a comparative benchmark, the analysis is repeated using the same dataset with nonparametric Gaussian process regression (GPR) model that does not incorporate physical information. Using MBDoE principles, we find that only five experiments are necessary to calibrate the nonlinear physics-informed parametric model, and with said limited data, this model outperforms the black-box machine learning GPR model. This key result underscores an emerging trend in the data science community: incorporating physical information into predictive models often drastically reduces the data requirements. Leveraging MBDoE further increased the data efficiency. By design, the proposed data science framework is general in nature and can be easily extended to other experimental and additive manufacturing systems beyond AJP.

Aerosol jet printing↗

GeoThermalCloud for EGS – An Open-source, User-friendly, Scalable AI Workflow for Modeling Enhanced Geothermal Systems

Enhanced Geothermal Systems (EGS) offer a vast potential to expand the use of geothermal energy. Heat is extracted from this engineered system by injecting relatively cold water into subsurface fractures, which are in contact with hot dry rock, and brought back to surface through production wells. Creating EGS requires improving the natural permeability of hot crystalline rocks. In this short conference paper, we present a reproducible workflow for modeling EGS. Our workflow called the GeoThermalCloud (GTC) for EGS, leverages recent advances in machine learning, deep learning, and high-performance computing. This GTC framework is currently being made open-source, user-friendly, and reproducible through python scripts as well as Google Colab/Jupyter Notebooks. This GTC for EGS modeling scripts are made available at https://github.com/SmartTensors/GeoThermalCloud.jl/tree/master/EGS and will constantly be updated to cater for geothermal community. Current GTC framework provides scripts to train deep learning (DL) models for techno-economics and data worth analysis. The Geothermal Design Tool (https://github.com/GeoDesignTool/GeoDT.git), a fast and simplified multi-physics solver, is used to develop a database for training DL models. This short paper provides details on the scripts to curate, process, and train DL models. The scripts can easily be modified to train on databases generated by other popular open-source simulators such as PFLOTRAN, STOMP, TOUGH, and GEOSX or commercial software such as ResFrac and COMSOL.

15 GEOTHERMAL ENERGY↗

Analysis of two-color photoelectron spectroscopy for attosecond metrology at seeded free-electron lasers

The generation of attosecond pulse trains at free-electron lasers opens new opportunities in ultrafast science, as it gives access, for the first time, to reproducible, programmable, extreme ultraviolet (XUV) waveforms with high intensity. In this work, we present a detailed analysis of the theoretical model underlying the temporal characterization of the attosecond pulse trains recently generated at the free-electron laser FERMI. In particular, the validity of the approximations used for the correlated analysis of the photoelectron spectra generated in the two-color photoionization experiments are thoroughly discussed. The ranges of validity of the assumptions, in connection with the main experimental parameters, are derived.

74 ATOMIC AND MOLECULAR PHYSICS↗

Line Faults Classification Using Machine Learning on Three Phase Voltages Extracted from Large Dataset of PMU Measurements

An end-to-end supervised learning method is developed to classify transmission line faults in a twoyear field-recorded dataset that includes synchronized measurements of three-phase voltages recorded by 38 Phasor Measurement Units (PMU) sparsely located in in the US Western Grid interconnection. Statistical analysis is performed to extract features from this large dataset to train Support Vector Machine (SVM), Random Forest (RF), and eXtreme Gradient Boosting (XGBoost) classifiers initially. The training further leverages a simulated dataset from a synthetic grid with 12 PMUs to increase the number of faults of types infrequently seen in the field-recorded dataset. Training the classification models with the combined dataset resulted in a classification accuracy of 97.7%. This is a significant improvement over 89.7% to 92.5% accuracy obtained by relying on the field-recorded dataset alone.

47 OTHER INSTRUMENTATION↗