Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “accelerator simulation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Geant4 Event Biasing and Fast Simulation

Geant4 offers advanced event biasing techniques to significantly accelerate simulations involving rare events. Various biasing methods, such as leading particle selection, cross-section biasing, radioactive decay enhancement, and bremsstrahlung splitting, enable efficient event sampling, though they require careful handling. Additionally, Geant4 provides a Fast Simulation Interface, allowing the replacement of standard processes in specific region and for selected particles, enabling faster execution or external code integration. Applications of fast simulation include electromagnetic shower modeling in calorimeters, machine learning inference, and offloading tasks to specialized hardware like GPUs, making Geant4 a powerful tool for computationally demanding simulations.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Linear Solvers for Collector Systems of Generalized Large-scale Inverter-Based Resources

Collector systems for inverter-based resources (IBRs) are typically represented by equivalent circuits for electromagnetic transient (EMT) simulations. Recent studies have revealed that modeling a detailed collector system is essential to accurately represent the behavior of IBRs, especially when dealing with partial tripping during external disturbances. However, there are several challenges in simulating a detailed EMT model of a collector system due to the time required to simulate such systems. Thus, this paper investigates the modeling of a detailed collector system, taking into account its configuration and components as defined in IEEE standard 2800. The configurations include the collector systems of generalized large-scale IBR plants. The components include the main IBR transformer, collector bus, and feeders with lines and/or cables. The EMT model of the collector system is represented by differential algebraic equations (DAEs) that are discretized to form linear equations that are solved using linear solvers. In this paper, linear solvers are proposed based on the Schur complement method, which are utilized for simulation of the EMT model of collector systems of generalized large-scale IBRs to accelerate simulation speed while maintaining the accuracy of the results. The proposed solvers are verified by comparing the performance to that of linear solvers provided in MATLAB.

Choi, Jongchan↗

Technical Report for Bayesian Optimization and Reinforcement Learning for Beam Polarization Increase in the BNL Hadron Injectors

This project developed and evaluated physics-informed Bayesian learning and machine learning (ML)-based optimization methods for improving beam polarization preservation in the BNL hadron injector chain. The work focused on uncertainty-aware digital twin modeling, Bayesian calibration of accelerator simulations using beam measurements, and data-efficient optimization strategies including Bayesian optimization and reinforcement learning. These methods were applied to injector tuning and RF control problems in realistic accelerator settings to support improved operational robustness and readiness for RHIC operations and future Electron–Ion Collider facilities. No subject inventions were disclosed under this award.

43 PARTICLE ACCELERATORS↗

Accelerating kinetic plasma simulations with machine-learning-generated initial conditions

Computational models of plasma technologies often solve for the system operating conditions by time-stepping an initial value problem to a quasi-steady solution. However, the strongly nonlinear and multi-timescale nature of plasma dynamics often necessitate millions, or even hundreds of millions, of steps to reach convergence, reducing the effectiveness of these simulations for computer-aided engineering. We consider acceleration of kinetic plasma simulations via data-driven machine-learning-generated initial conditions, which initialize the simulations close to their final quasi-steady-state, thereby reducing the number of steps to reach convergence. Three machine-learning models are developed to predict the density and ion kinetic profiles of capacitively coupled plasma discharges relevant to the microelectronics industry. The models are trained on kinetic simulations over a range of device operating frequencies and pressures. Best performance was observed when simulations were initialized with ion kinetic profiles generated by a convolutional neural network, reducing the mean number of steps to reach convergence by 17.1× when compared to initialization with a zero-dimensional global model. We also outline a workflow for continuous data-driven model improvement and simulation speedup, with the aim of generating sufficient data for full device digital twins.

Artificial neural networks↗

Surrogate modeling of Cellular-Potts agent-based models as a segmentation task using the U-Net neural network architecture

The Cellular-Potts model is a powerful and ubiquitous framework for developing computational models for simulating complex multicellular biological systems. Cellular-Potts models (CPMs) are often computationally expensive due to the explicit modeling of interactions among large numbers of individual model agents and diffusive fields described by partial differential equations (PDEs). In this work, we develop a convolutional neural network (CNN) surrogate model using a U-Net architecture that accounts for periodic boundary conditions. We use this model to accelerate the evaluation of a mechanistic CPM previously used to investigate in vitro vasculogenesis. The surrogate model was trained to predict 100 computational steps ahead (Monte-Carlo steps, MCS), accelerating simulation evaluations by a factor of 562 times compared to single-core CPM code execution on CPU. Over short timescales of up to 3 recursive evaluations, or 300 MCS, our model captures the emergent behaviors demonstrated by the original Cellular-Potts model such as vessel sprouting, extension and anastomosis, and contraction of vascular lacunae. This approach demonstrates the potential for deep learning to serve as a step toward efficient surrogate models for CPM simulations, enabling faster evaluation of computationally expensive CPM simulations of biological processes.

97 MATHEMATICS AND COMPUTING↗

Direct Acceleration of an Electron Beam with a Radially Polarized Long-Wave Infrared Laser

Direct laser acceleration with radially polarized lasers is an intriguing variant of laser-based particle acceleration that has the potential of offering GeV/cm-level energy while avoiding the instabilities and complex beam dynamics associated with plasma wakefield accelerators. A major limiting factor is the difficulty of generating high-power radially polarized beams. In this paper, we propose the use of CO2-based long-wave infrared (LWIR) lasers as a driver for direct laser acceleration, as the polarization insensitivity of the gain medium allows a radially polarized beam to be amplified. Additionally, the larger waist sizes, Rayleigh lengths, and pulse lengths associated with the long wavelength could improve the injection efficiency of the electron beam. By comparing acceleration simulations using a near-infrared laser and an LWIR laser, we show that the injection efficiency is indeed improved by up to an order of magnitude with the longer wavelength. Furthermore, we show that even sub-TW peak powers with an LWIR laser can provide MeV-level energy gains. Thus, radially polarized LWIR lasers show significant promise as a driver of a direct laser-driven demonstration accelerator.

43 PARTICLE ACCELERATORS↗

Parallel-in-Time Solution of Allen-Cahn Equations by Integrating Operator Learning into the Parareal Method

While recent advances in deep learning have shown promising efficiency gains in solving time-dependent partial differential equations (PDEs), matching the accuracy of conventional numerical solvers still remains a challenge. One strategy to improve the accuracy of deep learning-based solutions for time-dependent PDEs is to use the learned model as the coarse propagator in the Parareal method and a traditional numerical method as the fine solver. However, successful integration of deep learning into the Parareal method requires consistency between the coarse and fine solvers, particularly for PDEs exhibiting rapid changes such as sharp transitions. Here, to ensure this consistency, we propose using convolutional neural networks (CNNs) to learn the fully discrete time-stepping operator defined by the same numerical scheme employed as the fine solver. We demonstrate the effectiveness of the proposed method in solving the classical and mass-conservative Allen–Cahn (AC) equations. Through iterative updates in the Parareal algorithm, our approach achieves a significant computational speedup compared to traditional fine solvers while converging to high-accuracy solutions. Our results highlight that the proposed hybrid Parareal algorithm effectively accelerates simulations, particularly when implemented on multiple GPUs, and converges to the desired accuracy in only a few iterations. Another advantage of our method is that the CNN model is trained on trajectory-based data generated from random initial conditions, such that the trained model can be used to solve the AC equations with various initial conditions without retraining. This work demonstrates the potential of integrating neural network methods into parallel-in-time frameworks for efficient and accurate simulations of time-dependent PDEs.

97 MATHEMATICS AND COMPUTING↗

Dynamic Metal–Support Interaction Dictates Cu Nanoparticle Sintering on Al 2 O 3 Surfaces

Nanoparticle sintering remains a critical challenge in heterogeneous catalysis. In this work, we present a unified deep potential (DP) model based on the Perdew–Burke–Ernzerhof approximation of density functional theory for Cu nanoparticles on three Al 2 O 3 surfaces (γ-Al 2 O 3 (100), γ-Al 2 O 3 (110), and α-Al 2 O 3 (0001)). Using DP-accelerated simulations, we reveal that the nanoparticle size-mobility relationship strongly depends on the supporting surface. The diffusion of nanoparticles on the two γ-Al 2 O 3 surfaces is almost independent of the size of the nanoparticle, while the diffusion on α-Al 2 O 3 (0001) decreases rapidly with increasing size. Interestingly, nanoparticles with fewer than 55 atoms diffuse several times faster on α-Al 2 O 3 (0001) than on γ-Al 2 O 3 (100) at 800 K while expected to be more sluggish based on their larger binding energy at 0 K. The diffusion on α-Al 2 O 3 (0001) is facilitated by dynamic metal–support interaction (MSI), where Al atoms move out of the surface plane to optimize contact with the nanoparticle and relax back to the plane as the nanoparticle moves away. In contrast, the MSI on γ-Al 2 O 3 (100) and on γ-Al 2 O 3 (110) is dominated by more stable and directional Cu–O bonds, consistent with the limited diffusion observed on these surfaces. Our extended MD simulations provide insight into the sintering processes, showing that the dispersity of the nanoparticles strongly influences the coalescence driven by nanoparticle diffusion. We observed that the coalescence of Cu 13 nanoparticles on α-Al 2 O 3 (0001) can occur in a short time (10 ns) at 800 K even with an initial internanoparticle distance increased to 3 nm, while the coalescence on the two γ-Al 2 O 3 surfaces are inhibited significantly by increasing the initial internanoparticle distance. These findings demonstrate that the dynamics of the supporting surface is crucial to understanding the sintering mechanism and offer guidance for designing sinter-resistant catalysts by engineering the support morphology.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Seed-Mediated Colloidal Synthesis of Multimetallic and High-Entropy Alloy Nanocrystal Libraries with Enhanced Catalytic Performance

Engineering colloidally stable multimetallic nanocrystals offers many benefits in a wide range of applications and allows manipulation of physical, chemical, and electronic properties of materials at the nanoscale. Synthesis routes are challenged by the chemical complexity required to temporally and spatially coordinate the reduction and alloying of multiple metal species, which has hampered the development of tunable libraries of colloidal materials to date. Here, in this work, we demonstrate a seed-mediated synthesis method to incorporate five or more metal elements into uniform, colloidally stable nanocrystals. By integrating machine learning-accelerated simulations, the synthesis of shortlisted high-entropy alloy nanocrystals was demonstrated. Multiple seed materials can be used, leading to a library of multimetallic nanocrystals with tunable electronic, physical, and alloy structures. The advantage of this synthetic protocol is highlighted in the preparation of catalytic materials that showed 2 orders of magnitude higher reaction rates than monometallic catalysts and outstanding thermal stability, thus highlighting the promise of this approach for high-performance materials in many areas.

77 NANOSCIENCE AND NANOTECHNOLOGY↗

Diameter-dependent multiple proton jumps dictate hydronium and hydroxide transport in carbon nanotubes

Nanofluidic channels impose extreme confinement on water, giving rise to unusual transport phenomena of the liquid. However, how the transport of hydroxide and hydronium ions is influenced by such confinement is still not fully understood. This study employs machine learning-accelerated simulations, based on the SCAN density functional, to investigate proton transfer dynamics in CNTs of varying diameters (0.8 nm to 2.8 nm). The extreme confinement of water inside a 0.8 nm CNT not only enhances the probability of multiple consecutive proton jumps, but also reverses the relative diffusion coefficient of hydronium and hydroxide ions in bulk water. In CNTs with diameters larger than 0.8 nm, hydronium diffuses slightly faster than in bulk water, whereas hydroxide diffusion slows because of its localization near CNT walls, hindering multiple proton jumps. This work highlights the significant impact of nanoscale confinement on proton transfer dynamics, with implications for designing nanoscale systems with controlled proton transport.

Chemistry↗

Revealing EDL-driven reduction mechanisms in binary, ternary, and quaternary fluorinated electrolytes via an integrated MD–DFT–ML framework

Accurately predicting solid electrolyte interphase (SEI) formation requires explicitly resolving the electric double layer (EDL) structure, which deviates significantly from that of the bulk electrolyte. Although an established molecular dynamics (MD) and Density Functional Theory (DFT) framework can model SEI formation by evaluating reduction reactions of local clusters in the EDL, it suffers from a combinatorial computational bottleneck. To overcome this limitation, we introduce a machine-learning-accelerated simulation workflow (MD–DFT–ML), integrating a gradient-boosted regression model trained on EDL composition data to efficiently predict reduction potentials. We apply this framework to seven fluorinated electrolytes comprising fluorinated anions, a fluorinated ester solvent, two types of diluent (ion-solvating ester vs. non-solvating ether), and an FEC additive. The analysis shows that the EDL selectively accumulates cation-binding species; consequently, the non–cation-binding ether diluent rarely enters the EDL and makes minimal contributions to SEI formation. DFT calculations on statistically representative EDL clusters provide reduction potentials and fluorine-release pathways, while the ML model, which substantially reduces the DFT workload, predicts cluster reduction energies with a mean absolute error of 0.1 eV. The combined MD–DFT–ML approach also quantifies contributions from different sources to LiF formation in the SEI. This methodology establishes a generalizable route for multiscale modeling electrolyte and interphase design for next-generation electrochemical energy-storage systems.

DFT-MD-ML workflow↗

Efficient exact exchange using Wannier functions and other related developments in planewave-pseudopotential implementation of RT-TDDFT

The plane-wave pseudopotential (PW-PP) formalism is widely used for the first-principles electronic structure calculation of extended periodic systems. The PW-PP approach has also been adapted for real-time time-dependent density functional theory (RT-TDDFT) to investigate time-dependent electronic dynamical phenomena. In this work, we detail recent advances in the PW-PP formalism for RT-TDDFT, particularly how maximally localized Wannier functions (MLWFs) are used to accelerate simulations using the exact exchange. We also discuss several related developments, including an anti-Hermitian correction for the time-dependent MLWFs (TD-MLWFs) when a time-dependent electric field is applied, the refinement procedure for TD-MLWFs, comparison of the velocity and length gauge approaches for applying an electric field, and elimination of long-range electrostatic interaction, as well as usage of a complex absorbing potential for modeling isolated systems when using the PW-PP formalism.

Chemistry↗

aiida-flux-scheduler

AiiDA is a workflow management software that is capable of accelerating simulations on HPC machines. Currently, there is no scheduler plugin for flux. The current code that is being submitted to be released is the initial alpha version. The code will be hosted on the external LLNL github group.

Keilbart, Nathan [Lawrence Livermore National Labo↗

Evaluation of data driven low-rank matrix factorization for accelerated solutions of the Vlasov equation

Low-rank methods have shown success in accelerating simulations of a collisionless plasma described by the Vlasov equation, but still rely on computationally costly linear algebra every time step. We propose a data-driven factorization method using artificial neural networks, specifically with convolutional layer architecture, that trains on existing simulation data. At inference time, the model outputs a low-rank decomposition of the distribution field of the charged particles, and we demonstrate that this step is faster than the standard linear algebra technique. Numerical experiments show that the method achieves comparable reconstruction accuracy for interpolation tasks, generalizing to unseen test data in a manner beyond just memorizing training data; patterns in factorization also inherently followed the same numerical trend as those within algebraic methods (e.g., truncated singular-value decomposition). However, when training on the first 70% of a time-series data and testing on the remaining 30%, the method fails to meaningfully extrapolate. Despite this limiting result, the technique may have benefits for simulations in a statistical steady-state or otherwise showing temporal stability. These results suggest that while the model offers a computationally efficient alternative for datasets with temporal stability, its current formulation is best suited for interpolation rather than for predicting future states in time-evolving systems. This study thus lays the groundwork for further refinement of neural network-based approaches to low-rank matrix factorization in high-dimensional plasma simulations.

97 MATHEMATICS AND COMPUTING↗

Accelerating resonant spectroscopy simulations using multishifted biconjugate gradient

Resonant spectroscopies, which involve intermediate states with finite lifetimes, provide important insights into collective excitations in quantum materials that are otherwise inaccessible. However, theoretical understanding in this area is often limited by the numerical challenges of solving Kramers-Heisenberg-type response functions for large-scale systems. To address this, we introduce a multishifted biconjugate gradient algorithm that exploits the shared structure of Krylov subspaces across spectra with varying incident energies, effectively reducing the computational complexity to that of linear spectroscopies. Both mathematical proofs and numerical benchmarks confirm that this algorithm substantially accelerates spectral simulations, achieving constant complexity independent of the number of incident energies, while ensuring accuracy and stability. This development provides a scalable, versatile framework for simulating advanced spectroscopies in quantum materials.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Unveiling the lithium-ion transport mechanism in Li{sub 2}ZrCl{sub 6} Solid-State Electrolyte {ital via} deep learning-accelerated molecular dynamics simulations.

Lithium zirconium chlorides (LZCs) present a promising class of cost-effective solid electrolytes for next-generation all-solid-state batteries. The unique crystal structure of LZCs plays a crucial role in facilitating lithium-ion mobility, which further affects the electrochemical performance. To understand the underlying mechanism governing ion transport, we employed deep learning-accelerated molecular dynamics simulation on Li2ZrCl6 (trigonal alpha- and monoclinic beta-LZC), focusing specifically on the zirconium coordination environment. Our results reveal that disordered alpha-LZC exhibits the highest ionic conductivity, while beta-LZC demonstrates significantly lower conductivity, closely aligning with experimental findings. The study confirms that across all phases, lithium migration proceeds via the site-to-site hopping mechanism, where variations in site residence times critically impact the overall ionic conductivity. In alpha-LZCs, lithium ions prefer to anisotropically diffuse across interlayers as the result of a lower energy barrier, driven primarily by collective diffusion. In contrast, lithium ions in beta-LZC primarily isotropically diffuse within the intralayer, hindered by higher energy barriers and determined by individual diffusion. The variation in ZrCl6 2- octahedral unit softening, induced by the specific layered arrangement of zirconium atoms, emerges as a critical determinant of the energy barriers across the LZC phases. These atomic-scale insights into the transport processes provide valuable guidance for the rational design and optimization of LZCs-based electrolytes, accelerating their practical application in advanced energy storage technologies.

Guo, Hanzeng↗

Parallel quantum computing simulations via quantum accelerator platform virtualization

Quantum circuit execution is a central task in quantum computation. Due to inherent quantum-mechanical constraints, quantum computing workflows often involve a considerable number of independent measurements over a large set of slightly different quantum circuits. Here we discuss a simple model for parallelizing such quantum circuit executions that is based on introducing a large array of virtual quantum processing units (mapped to HPC nodes in our case) as a parallel quantum computing platform. Implemented within the XACC framework, the model can readily take advantage of its backend-agnostic features, enabling parallel quantum computing/simulation over any target backend supported by XACC. We illustrate the performance of this approach by demonstrating strong scaling in two pertinent domain science problems, namely in computing the gradients for the multi-contracted variational quantum eigensolver and in data-driven quantum circuit learning, where we vary the number of qubits and the number of circuit layers. Here, the latter simulation leverages the cuQuantum library to run efficiently on GPU-accelerated HPC platforms.

97 MATHEMATICS AND COMPUTING↗

Time-dependent-bases with local CUR decomposition method for accelerating turbulent combustion simulations

Here, this study presents a novel reduced-order modeling framework, Time-Dependent Bases with Local CUR decomposition (TDB-L-CUR), designed to efficiently and accurately approximate the species transport equations in reacting flow simulations. The method extends the existing TDB-CUR approach for chemically reacting flows (Jung et al. Comput. Methods Appl. Mech. Engrg. 437 (2025) 117758), which leverages matrix decomposition techniques to form a global-in-space, time-dependent low-dimensional manifold. While TDB-CUR performs well in homogeneous systems, it may be less well-suited to spatially heterogeneous systems such as turbulent flames, where higher-rank approximations are typically required. The proposed TDB-L-CUR framework introduces two methodological extensions to the baseline approach. First, it applies unsupervised clustering to partition the physical domain into distinct regions, enabling spatially localized manifold construction, thereby reducing the rank required for the reduced-order representation. Second, it incorporates a computational singular perturbation (CSP)-based scheme for identifying and penalizing fast species, allowing for spatio-temporally adaptive mitigation of chemical stiffness. The proposed framework is validated on a hierarchy of test cases, including a one-dimensional premixed flame, a two-dimensional nonpremixed ignition case with vortex interaction, and a three-dimensional turbulent premixed flame. TDB-L-CUR significantly improves accuracy over TDB-CUR while further reducing computational cost. The fully on-the-fly formulation of TDB-L-CUR (i.e., requiring no offline training or prior knowledge) makes it a robust and scalable tool for reduced-order modeling of reactive flows.

Local manifold↗