Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “training”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

High temperature melting of dense molecular hydrogen from machine-learning interatomic potentials trained on quantum Monte Carlo

We present results and discuss methods for computing the melting temperature of dense molecular hydrogen using a machine learned model trained on quantum Monte Carlo data. In this newly trained model, we emphasize the importance of accurate total energies in the training. We integrate a two phase method for estimating the melting temperature with estimates from the Clausius–Clapeyron relation to provide a more accurate melting curve from the model. We make detailed predictions of the melting temperature, solid and liquid volumes, latent heat, and internal energy from 50 to 180 GPa for both classical hydrogen and quantum hydrogen. At pressures of roughly 173 GPa and 1635 K, we observe molecular dissociation in the liquid phase. Here, we compare with previous simulations and experimental measurements.

08 HYDROGEN↗

Industry-driven Training and Curriculum Development Process

The development of a sustainable, skilled fusion workforce requires coordinated strategy between all sectors of fusion industry. This paper outlines a framework to align training programs with evolving technical and professional demands of fusion, including enhancing existing curricula, the establishment of new programs at educational institutions, and the identification of workforce gaps informed through industry engagement. Effective curriculum development requires input from both educators and employers to ensure that academic content reflects real-world challenges and can prepare students for successful transitions into the field. Collaborative models, such as industry-led training programs, inter-institutional partnerships, and faculty development initiatives, are highlighted as mechanisms for scalable and inclusive workforce development. Continued program success and relevance will be dependent on continuous review processes, including feedback from employers, alumni, and advisory boards. The combination of these programs supports the formation of flexible, industry-informed training pathways. This approach aims to foster a competent workforce capable of advancing fusion energy research and commercialization.

Gehrig, Monica [ORNL] (ORCID:0000000341022612)↗

Training quantum neural networks using the quantum information bottleneck method

Abstract We provide in this paper a concrete method for training a quantum neural network to maximize the relevant information about a property that is transmitted through the network. This is significant because it gives an operationally well founded quantity to optimize when training autoencoders for problems where the inputs and outputs are fully quantum. We provide a rigorous algorithm for computing the value of the quantum information bottleneck quantity within error ε that requires O ( log 2 ⁡ ( 1 / ϵ ) + 1 / δ 2 ) queries to a purification of the input density operator if its spectrum is supported on { 0 } ⋃ [ δ , 1 − δ ] for δ > 0 and the kernels of the relevant density matrices are disjoint. We further provide algorithms for estimating the derivatives of the QIB function, showing that quantum neural networks can be trained efficiently using the QIB quantity given that the number of gradient steps required is polynomial.

Çatlı, Ahmet Burak (ORCID:0000000152294141)↗

Breaking the barrier of human-annotated training data for machine learning-aided plant research using aerial imagery

Machine learning (ML) can accelerate biological research. However, the adoption of such tools to facilitate phenotyping based on sensor data has been limited by (i) the need for a large amount of human-annotated training data for each context in which the tool is used and (ii) phenotypes varying across contexts defined in terms of genetics and environment. This is a major bottleneck because acquiring training data is generally costly and time-consuming. This study demonstrates how a ML approach can address these challenges by minimizing the amount of human supervision needed for tool building. A case study was performed to compare ML approaches that examine images collected by an uncrewed aerial vehicle to determine the presence/absence of panicles (i.e. “heading”) across thousands of field plots containing genetically diverse breeding populations of 2 Miscanthus species. Automated analysis of aerial imagery enabled the identification of heading approximately 9 times faster than in-field visual inspection by humans. Leveraging an Efficiently Supervised Generative Adversarial Network (ESGAN) learning strategy reduced the requirement for human-annotated data by 1 to 2 orders of magnitude compared to traditional, fully supervised learning approaches. The ESGAN model learned the salient features of the data set by using thousands of unlabeled images to inform the discriminative ability of a classifier so that it required minimal human-labeled training data. This method can accelerate the phenotyping of heading date as a measure of flowering time in Miscanthus across diverse contexts (e.g. in multistate trials) and opens avenues to promote the broad adoption of ML tools.

59 BASIC BIOLOGICAL SCIENCES↗

Hamiltonian learning using machine-learning models trained with continuous measurements

Here, we build upon recent work on the use of machine-learning models to estimate Hamiltonian parameters using continuous weak measurement of qubits as input. We consider two settings for the training of our model: (1) supervised learning, where the weak-measurement training record can be labeled with known Hamiltonian parameters, and (2) unsupervised learning, where no labels are available. The first has the advantage of not requiring an explicit representation of the quantum state, thus potentially scaling very favorably to a larger number of qubits. The second requires the implementation of a physical model to map the Hamiltonian parameters to a measurement record, which we implement using an integrator of the physical model with a recurrent neural network to provide a model-free correction at every time step to account for small effects not captured by the physical model. We test our construction on a system of two qubits and demonstrate accurate prediction of multiple physical parameters in both the supervised context and the unsupervised context. We demonstrate that the model benefits from larger training sets, establishing that it is “learning,” and we show robustness regarding errors in the assumed physical model by achieving accurate parameter estimation in the presence of unanticipated single-particle relaxation.

97 MATHEMATICS AND COMPUTING↗

Multiport Converter based Auxiliary Power Supply for Heavy Duty Fuel Cell Power Train

This paper presents a multiport converter (MPC) based power supply to charge the 12 V and 24 V auxiliary batteries in heavy duty (HD) fuel cell (FC) electric vehicle (EV) power train. Compared to conventional auxiliary power supply architectures, the proposed architecture shares the power electronic components of the main power train yet preserves the isolation and all other desirable features of auxiliary power supplies. The proposed auxiliary power supply architecture reduces the requirements of the devices, gate drivers and thermal requirements thereby having potential benefits towards improving the power density, cost and weight of the overall power train. The topology, design parameters and control methodology are discussed, and results are presented to validate the proposed architecture.

Mukherjee, Subho↗

MDLoader: A Hybrid Model-Driven Data Loader for Distributed Graph Neural Network Training

Scalable data management is essential for processing large scientific dataset on HPC platforms for distributed deep learning. In-memory distributed storage is preferred for its speed, enabling rapid, random, and frequent data access required by stochastic optimizers. Processes use one-sided or collective communication to fetch remote data, with optimal performance depending on (i) dataset characteristics, (ii) training scale, and (iii) interconnection network. Empirical analysis shows collective communication excels with larger mini-batch sizes and/or fewer processes, whereas one-sided communication outperforms at larger scales. We propose MDLoader, a hybrid in-memory data loader for distributed graph neural network training. MDLoader features a model-driven performance estimator that dynamically selects between one-sided and collective communication at the beginning of training using Tree of Parzen Estimators (TPE). Evaluations on NERSC Perlmutter and OLCF Summit show MDLoader outperforms single-backend loaders by up to 2.83 × and predicts the suitable communication method with 96.3% (Perlmutter) and 94.3% (Summit) success rate.

Bae, Jonghyun↗

Development of systematic uncertainty-aware neural network trainings for binned-likelihood analyses at the LHC

We propose a neural network training method capable of accounting for the effects of systematic variations of the data model in the training process and describe its extension towards neural network multiclass classification. The procedure is evaluated on the realistic case of the measurement of Higgs boson production via gluon fusion and vector boson fusion in the τ τ decay channel at the CMS experiment. The neural network output functions are used to infer the signal strengths for inclusive production of Higgs bosons as well as for their production via gluon fusion and vector boson fusion. We observe improvements of 12 and 16% in the uncertainty in the signal strengths for gluon and vector-boson fusion, respectively, compared with a conventional neural network training based on cross-entropy.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Intelligent Sampling of Extreme-Scale Turbulence Datasets for Accurate and Efficient Spatiotemporal Model Training

With the end of Moore’s law and Dennard scaling, efficient training increasingly requires rethinking data volume. Can we train better models with significantly less data via intelligent subsampling? To explore this, we develop SICKLE, a sparse intelligent curation framework for efficient learning, featuring a novel maximum entropy (MaxEnt) sampling approach, scalable training, and energy benchmarking. We compare MaxEnt with random and phase-space sampling on large direct numerical simulation (DNS) datasets of turbulence. Evaluating SICKLE at scale on Frontier, we show that subsampling as a preprocessing step can, in many cases, improve model accuracy and substantially lower energy consumption, with observed reductions of up to 38×.

Brewer, Wes [ORNL] (ORCID:0000000236393956)↗

Training NuGraph2 for ICARUS

This presentation describes the process of training NuGraph2, a Graphical Neural Network for event reconstruction, on simulated ICARUS neutrino event data. This began with an investigation into filtering ICARUS spacepoint data. Then NuGraph2 was repeatedly trained on three event samples, which were used for finding optimized machine-learning parameters and to find and fix the causes of several crashes in NuGraph2 s preprocessing and training scripts.

43 PARTICLE ACCELERATORS↗

Solar Training and Education Partnership for Underserved Populations

Solar Landscape’s STEP-UP program provided high quality solar installation training in partnership with community-based organizations (CBOs) in various regions within the U.S. Solar Landscape leveraged internal subject matter experts (SME’s) industry guidance and regional training assessments to provide customized training designed to support the growing solar and broader energy sector. Upon completion of the program, the team assisted trainees and nonprofit partners with connections to Solar Landscape contractors as well as local and National residential solar installation companies to facilitate placement into careers and apprenticeships.

14 SOLAR ENERGY↗

Stacked networks improve physics-informed training: Applications to neural networks and deep operator networks

Physics-informed neural networks and operator networks have shown promise for effectively solving equations modeling physical systems. However, these networks can happen to be difficult or impossible to train accurately. Here, we present a novel multifidelity framework for stacking physics-informed neural networks and operator networks that facilitates training. We successively build a chain of networks, where the output at one step can act as a low-fidelity input for training a longer chain, gradually increasing the expressivity of the learnt model. The equations imposed at each step of the iterative process can be the same or different (akin to simulated annealing). The iterative (stacking) nature of the proposed method allows us to learn progressively features of a solution which could have been hard to learn directly. Through benchmark problems including a nonlinear pendulum, the wave equation, and the viscous Burgers equation, we show how stacking can be used to improve the accuracy and reduce the required size of physics-informed neural networks and operator networks.

97 MATHEMATICS AND COMPUTING↗

Smart Building Technology Training Modules for Academic and Professional Education

Smart building technologies are a new suite of resources that improve building energy efficiency and resilience, reduce carbon emissions, and provide load flexibility to the grid. However, in both college curricula and building professionals’ continuing education, there is a lack of systematic instruction on smart building technologies–topics that include smart building concepts, key components, smart building controls, “Internet of Things” (IoT) devices, and how to integrate multiple energy systems including distributed energy resources (DER). This major gap in smart building education prevents stakeholders from understanding and adopting smart building technologies in building design and operations. Slipstream leads a DOE-funded project developing a semester-long smart building curriculum for college students and adapting the contents into 16 training videos for building professionals and the general public. The education and training cover the drivers and benefits of smart building technologies, key building energy systems, the latest sensor technologies and IoT devices, and focus on topics related to smart building controls (i.e., energy management information systems, smart building control platforms, cybersecurity, grid-interactive-efficient buildings (GEBs), smart building control methods, and occupant-centric control. This paper describes the project approach, provides outlines of the training materials, and identifies lessons learned in creating the content. We also suggest ways to scale the instruction of smart building concepts to empower the workforce to accelerate the adoption of smart building technologies in the real world.

99 GENERAL AND MISCELLANEOUS↗

Challenges in Training PINNs: A Loss Landscape Perspective

This paper explores challenges in training Physics Informed Neural Networks (PINNs), emphasizing the role of the loss landscape in the training process. We examine difficulties in minimizing the PINN loss function, particularly due to ill conditioning caused by differential operators in the residual term. We compare gradient-based optimizers Adam, L-BFGS, and their combination Adam+L-FGS, showing the superiority of Adam+L-BFGS, and introduce a novel secondorder optimizer, NysNewton-CG (NNCG), which significantly improves PINN performance. Theoretically, our work elucidates the connection between ill-conditioned differential operators and ill-conditioning in the PINN loss and shows the benefits of combining first- and second-order optimization methods. Our work presents valuable insights and more powerful optimization strategies for training PINNs, which could improve the utility of PINNs for solving difficult partial differential equations.

Rathore, Pratik↗

Lessons learned from the development and implementation of a workforce training curriculum for advanced controls for high performance HVAC systems

Over the past decade, academic research on advanced controls has slowly transitioned into new software platforms, giving rise to various companies developing and deploying these innovative products, including solutions for light commercial HVAC systems. However, the current workforce remains widely unprepared to install, maintain and operate these systems, particularly complex software-based control platforms, as most workforce training programs still focus on traditional building automation for large commercial buildings. This paper presents the development and piloting of curriculum for three key types of professionals: ● Technicians (trade-level): installing and maintaining modern high-performance HVAC systems and controls ● Programmers (undergrad-level): developing and implementing advanced controls ● Engineers and energy professionals (undergrad/grad-level): managing and evaluating system performance We share details of the material developed including training videos, open-source software, instruction manuals. We also present the results of a pilot implementation of the training materials with real students.

Casillas, Armando↗

Operation of helium sub-atmospheric multistage cryogenic centrifugal compressor trains: Part 1 – Steady state modeling and speed ratio selection

Helium cryogenic systems which can provide cooling below the normal boiling point of helium (approximately 4.2 K) are often required by superconducting radio-frequency niobium resonators utilized in modern high-energy particle accelerators. Achieving temperatures below 4.2 K generally involves operating a cryogenic vessel with liquid helium under sub-atmospheric conditions, thereby lowering the saturation pressure and corresponding saturation temperature. Over the last several decades, multi-stage cryogenic centrifugal compressor trains (CC’s) have been operated efficiently and reliably within large-scale cryogenic systems to continuously evacuate helium vapor generated by a device within the vessel, maintaining sub-atmospheric conditions in the vessel while pressurizing the return vapor to above atmospheric conditions. Traditionally, these CC systems have been operated using empirically derived control philosophies and insight gathered from previous operational experience. Recent efforts at the Facility for Rare Isotope Beams (FRIB) have been aimed at the development of a theoretical basis to characterize the operation of multi-stage cryogenic centrifugal compressor train and utilizing predictive model results to generate control parameters. The objective of this research was identifying operational points which adequately balance cryogenic system efficiency, stability, and overall ease of operation. Furthermore, this manuscript provides an overview of the predictive model development, characterization of the FRIB cryogenic centrifugal compressors and implementation of the predicted performance results during steady-state system operation.

Compressor train control↗

Rank-Limiting Strategies for Optimizing Tensor-Train Finite-Difference Time-Domain Simulations

We introduce rank-limiting strategies to optimize tensor-train decompositions for three-dimensional finite-difference time-domain simulations using the relationship between the tensors and their specific dimensionality. These include the use of hard caps on the inner ranks of the tensor train decomposition and the use of a group rounding algorithm taking into account all field components simultaneously. Here, several numerical examples are considered to verify the efficacy of the proposed optimization strategies.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

From Sim to Real: A Pipeline for Training and Deploying Traffic Smoothing Cruise Controllers

Designing and validating controllers for connected and automated vehicles to enhance traffic flow presents significant challenges, from the complexity of replicating real-world stop-and-go traffic dynamics in simulation, to the intricacies involved in transitioning from simulation to actual deployment. In this work, we present a full pipeline from data collection to controller deployment. Specifically, we collect 772 km of driving data from the I-24 in Tennessee, and use it to build a one-lane simulator, placing simulated vehicles behind real-world trajectories. Using policy-gradient methods with an asymmetric critic, we improve fuel efficiency by over 10% when simulating congested scenarios. Our comprehensive approach includes reinforcement learning for controller training, software verification, hardware validation and setup, and navigating various sim-to-real challenges. Furthermore, we analyze the controller's behavior and wave-smoothing properties, and deploy it on four Toyota Rav4’s in a real-world validation experiment on the I-24. Lastly, we release the driving dataset, the simulator and the trained controller, to enable future benchmarking and controller design.

42 ENGINEERING↗