Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “fast algorithm”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 235 records · Page 13

A Solution Method for the Filtered Lifting Line Theory

The filtered lifting line theory presents a continuous form of the inviscid momentum equations of flow over a lifting device, such as a wing or rotor blade, using body forces without mathematical singularities. This theory is also consistent with an actuator line representation of a lifting device. In this work, we present a reformulation of the equations in terms of the local flow angle along the line, which allows solving the stand-alone equations using multivariate root-finding algorithms. This approach can be used to obtain a fast, computationally inexpensive solution of the loading distribution along a wing without the need to perform computational fluid dynamic simulations. We study the requirements in terms of resolution in the spanwise direction and establish the criteria for spacing and minimum amount of points required along the blade to obtain converged solutions. The solutions are compared to results from large-eddy simulations, and we observed excellent agreement with less than a percent difference in quantities along the blade between the methods.

17 WIND ENERGY↗

Multiparticle cumulant mapping for Coulomb explosion imaging: Calculations and algorithm

We present a versatile cumulant mapping algorithm for analyzing correlated particle emission, offering insights into complex electronic and nuclear dynamics. Recently, we have demonstrated the use of cumulant mapping to extract information-rich correlations between the momenta of multiple fragments produced in Coulomb explosion imaging experiments [C. Cheng et al., Phys. Rev. Lett. 130, 093001 (2023)]. We define cumulant mapping in terms of histograms, enabling fast computation of linear (additive) observables. However, applying the same algorithm to nonlinear (nonadditive) observables poses challenges, as the computation time of conventional estimators scales nonlinearly with data size. To overcome this, we develop estimators and an accompanying algorithm to enable computationally efficient estimation of the cumulant of interest. Comparisons of computation times and signal-to-noise ratios reveal the superior performance of our approach. This method is demonstrated on the (D+, D+, C+, O+) dissociation channel of CD 2 ⁢O 4+ produced in a strong-field ionization experiment. Additionally, Poisson statistics are used to simulate the two methods and provide insights into the efficiency of our algorithm. The proposed methodology unlocks efficient computation of cumulant mapping for a broader range of complex systems and observables, such as the laser pulse dependence of ionization dynamics.

74 ATOMIC AND MOLECULAR PHYSICS↗

Developing a Consistent Travel-Time Framework for Comparing Three-Dimensional Velocity Models for Seismic Location Accuracy

Abstract Location algorithms have historically relied on simple, one-dimensional (1D) velocity models for fast seismic event locations. 1D models are generally used as travel-time lookup tables, one for each seismic phase, with travel-times pre-calculated for event distance and depth. These travel-time lookup tables are extremely fast to use and this fast computational speed makes them the preferred type of velocity model for operational needs. Higher-dimensional (i.e., three-dimensional—3D) seismic velocity models are becoming readily available and provide more accurate event locations over 1D models. The computational requirements of these 3D models tend to make their operational use prohibitive. Additionally, comparing location accuracy for 3D seismic velocity models tends to be problematic, as each model is determined using different ray-tracing algorithms. Attempting to use a different algorithm than the one used to develop a model usually results in poor travel-time prediction. We demonstrate and test a framework to create first-P and first-S 3D travel-time correction surfaces using an open-source framework ( PCalc + GeoTess , https://www.sandia.gov/salsa3d/software/geotess ) that easily stores 3D travel-time and uncertainty data. This framework produces fast travel-time and uncertainty predictions and overcomes the ray-tracing algorithm hurdle because the lookup tables can be generated using the exact ray-tracing algorithm that is preferred for a model.

3D velocity models↗

A Pulsar-Inspired Timing Framework for Power System: Optimization and Performance Evaluation

Due to their excellent stability, neutron pulsar stars are considered promising candidate timing sources for power system applications. However, the complexity of pulsar signals necessitates advanced processing algorithms to provide accurate timing references. This paper presents the foundational framework for pulsar signal processing, serving as the basis for further optimization. To enhance the timing accuracy and computation efficiency in pulsar period searches, three algorithms are proposed as the initial optimization step: wavelet de-noising, fast folding, and cross-correlation for profile evaluation. Wavelet de-noising improves signal-to-noise ratio (SNR) by 36%–70%. Fast folding reduces computation time from hundreds of seconds to mere milliseconds. Cross-correlation works better than traditional SNR-based methods by effectively identifying the optimal period. The performance of the proposed algorithms is evaluated using observation data from telescopes. Together, these algorithms significantly improve pulsar timing performance, reducing the error of the Pulse Per Second (PPS) signal from hundreds to tens of microseconds.

Wu, Ori [ORNL] (ORCID:0000000326723410)↗

Computational thermomechanics for crystalline rock. Part II: Chemo-damage-plasticity and healing in strongly anisotropic polycrystals

We present a thermal–mechanical–chemical-phase field model that captures the multi-physical coupling effects of precipitation creeping, crystal plasticity, anisotropic fracture, and crack healing in polycrystalline rock at various temperature and strain-rate regimes. This model is solved via a fast Fourier transfer solver with an operator-split algorithm to update displacement, temperature and phase field, and chemical concentration incrementally. In nuclear waste disposal in salt formation, brine inside the crystal salt may migrate along the grain boundary and cracks due to the gradient of interfacial energy and pressure. This migration has a significant implication on the permeability evolution, creep deformation, and crack healing within rock salt but is difficult to incorporate implicitly via effective medium theories compared with computational homogenization. As such, we introduce a thermodynamic framework and a corresponding computational implementation that explicitly captures the brine diffusion along the grain boundary and crack at the grain scale. Meanwhile, the anisotropic fracture and healing are captured via a high-order phase field that represents the regularized crack region in which a newly derived non-monotonic driving force is used to capture the fracture and healing due to the solution–precipitation. Numerical examples are presented to demonstrate the capacity of the thermodynamic framework to capture the multiphysics material behaviors of rock salt.

42 ENGINEERING↗

Efficient Treatment of Large Active Spaces through Multi-GPU Parallel Implementation of Direct Configuration Interaction

In this study, we have extended our graphical processing unit (GPU)-accelerated direct configuration interaction program to multiple devices, reducing iteration times for configuration spaces of 165 million determinants to only 3 s using NVIDIA P100 GPUs. Similar improvements in the one- and two-particle reduced density matrix formation allow for fast analytical energy gradients and electronic properties. Our parallel algorithm enables the calculation of arbitrarily large configuration spaces (limited only by available system memory), with iteration times of 13 min for an active space of 18 electrons in 18 orbitals (2.4 billion determinants) using six consumer grade NVIDIA 1080Ti GPUs. These advances enable routine molecular dynamics simulations, geometry optimizations, and absorption spectrum calculations for molecules with large configuration spaces, a task that has heretofore required massive computational effort. In this work, we demonstrate the utility of our program by generating the absorption spectrum for diphenyl acetylene at the floating occupation molecular orbital complete active space configuration interaction level of theory. Lastly, several active spaces were investigated to assess the dependence of spectral features on orbital space dimension.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Conceptual study of a two-layer silicon pixel detector to tag the passage of muons from cosmic sources through quantum processors

Abstract Recent studies in quantum computing have shown that quantum error correction with large numbers of physical qubits are limited by ionizing radiation from high-energy particles. Depending on the physical setup of the quantum processor, the contribution of muons from cosmic sources can constitute a significant fraction of these interactions. As most of these muons are difficult to stop, we perform a conceptual study of a two-layer silicon pixel detector to tag their hits on a solid-state quantum processor instead. With a typical dilution refrigerator geometry model, we find that efficiencies greater than 50% are most likely to be achieved if at least one of the layers is operated at the deep-cryogenic (<1 K) flanges of the refrigerator. Following this finding, we further propose a novel research program that could allow the development of silicon pixel detectors that are fast enough to provide input to quantum error correction algorithms, can operate at deep-cryogenic temperatures, and have very low power consumption.

Instruments & Instrumentation↗

Improving reproducibility in synchrotron tomography using implementation-adapted filters

For reconstructing large tomographic datasets fast, filtered backprojection-type or Fourier-based algorithms are still the method of choice, as they have been for decades. These robust and computationally efficient algorithms have been integrated in a broad range of software packages. The continuous mathematical formulas used for image reconstruction in such algorithms are unambiguous. However, variations in discretization and interpolation result in quantitative differences between reconstructed images, and corresponding segmentations, obtained from different software. This hinders reproducibility of experimental results, making it difficult to ensure that results and conclusions from experiments can be reproduced at different facilities or using different software. In this paper, a way to reduce such differences by optimizing the filter used in analytical algorithms is proposed. These filters can be computed using a wrapper routine around a black-box implementation of a reconstruction algorithm, and lead to quantitatively similar reconstructions. Use cases for this approach are demonstrated by computing implementation-adapted filters for several open-source implementations and applying them to simulated phantoms and real-world data acquired at the synchrotron. Our contribution to a reproducible reconstruction step forms a building block towards a fully reproducible synchrotron tomography data processing pipeline.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Collaborative Fault Tolerant Control of Non-Signalized Intersections for Connected and Autonomous Vehicles

With the potential of increased penetration of connected and autonomous vehicles (CAVs), intersectional signal control faces new challenges in terms of its operation and implementation. One possibility is to fully make use of the communication capabilities of CAVs so that intersectional signal control can be realized by CAVs alone – this leads to non-signalized intersectional operation for traffic networks in urban areas. In this paper, the state-of-the-art on collaborative fault tolerant control schemes for complex systems will be briefly described. This is then followed by the formulation of operational fault tolerant control that realizes the collaborative fault tolerance functionality at CAVs operational level in response to possible individual vehicle faults, where detailed modelling using vehicle movement dynamics will be described together with the construction of fast fault diagnosis and a collaborative fault tolerant control algorithm. A simple example will be given as well to demonstrate the proposed algorithm together with the discussions on other issues such as randomness of the system, communication errors and full energy consideration. These leads to several future directions of the research for the traffic flow control of non-signalized intersections with 100% penetration of CAVs.

Wang, Hong↗

Simplified Transactive Distribution Grids for Bulk Power System Mechanism Development

As distributed energy resources and smart devices become omnipresent in the electrical power grid, transactive energy control mechanisms are evolving. From real-time to day-ahead markets, these transactive energy algorithms involve more and more agents, whose behavior is going to affect the transmission and generation network. In order to study the interaction between the wholesale and retail energy markets, extensive co-simulations are performed. To be able to redesign, evaluate, and verify new control algorithms, the simulations need to provide results in a fast and reliable manner. This work has built and tested a transactive distribution grid model, the DSO-Stub, meant to offer a configurable distribution retail market while ensuring the computational burden is not significantly increased.

Transactive Energy, , Distribution System Operator↗

Maximum Power Reference Tracking Algorithm for Power Curtailment of Photovoltaic Systems

This paper presents an algorithm for power curtailment of photovoltaic (PV) systems under fast solar irradiance intermittency. Based on the Perturb and Observe (P&O) technique, the method contains an adaptive gain that is compensated in real-time to account for moments of lower power availability. In addition, an accumulator is added to the calculation of the step size to reduce the overshoot caused by large irradiance swings. A testbed of a three-phase single-stage, 500 kVA PV system is developed on the OPAL-RT eMEGAsim real-time simulator. Field irradiance data and a regulation signal from PJM (RTO) are used to compare the performance of the proposed method with other techniques found in the literature. Results indicate an operation with smaller overshoot, less dc-link voltage oscillations, and improved power reference tracking capability.

Paduani, Victor↗

A Fast VANET-Assisted Scheme for Event Data Recorders

An event data recorder (EDR) is a device installed in a vehicle to record information. Similar to a black box in an airplane, an EDR is used in the study of automobile accidents. Many schemes have been proposed that use vehicle network technology to help record EDR data, including schemes involving storing data on roadside units or nearby vehicles and schemes leveraging blockchain technology. However, these schemes do not take into account the vehicle company’s server; with the increased use of autonomous vehicles, the data related to these vehicles are always uploaded to the vehicle company’s server. In this scenario, we classify the situation into different cases, according to whether or not it is an emergency and whether the vehicle and the server are connected. For these cases, we propose a scheme whereby a vehicle uploads the EDR data to a cloud server and sends the evidence of storage to the nearby vehicle through a vehicular ad hoc network. Our scheme offers a fast response due to the use of symmetric cryptography algorithms while also considering security requirements.

Liu, Wei↗

3D diffractive imaging of nanoparticle ensembles using an x-ray laser

Single particle imaging at x-ray free electron lasers (XFELs) has the potential to determine the structure and dynamics of single biomolecules at room temperature. Two major hurdles have prevented this potential from being reached, namely, the collection of sufficient high-quality diffraction patterns and robust computational purification to overcome structural heterogeneity. We report the breaking of both of these barriers using gold nanoparticle test samples, recording around 10 million diffraction patterns at the European XFEL and structurally and orientationally sorting the patterns to obtain better than 3-nm-resolution 3D reconstructions for each of four samples. With these new developments, integrating advancements in x-ray sources, fast-framing detectors, efficient sample delivery, and data analysis algorithms, we illuminate the path towards sub-nanometer biomolecular imaging. The methods developed here can also be extended to characterize ensembles that are inherently diverse to obtain their full structural landscape.

47 OTHER INSTRUMENTATION↗

Results and lessons learned from accelerating radio frequency modeling using machine learning [slides]

The “advanced tokamak” reactor concept is a leading candidate for a steady state fusion pilot plant. An advanced tokamak (AT) sustains a majority of the required plasma current with effects resulting from maintenance of the peaked pressure at the device center. This current is augmented by auxiliary current drive sources. These auxiliary actuators may consist of neutral particle beams and/or radio frequency (RF) systems such as lower hybrid current drive (LHCD) and high harmonic fast wave (HHFW) current drive using radio and microwaves from antennas. The primary focus of this work is to develop models of RF current profile control suitable for use in integrated modeling frameworks and for real-time control in experiments. Direct physics models of RF current drive can be computationally intensive. In order to achieve predictive times appropriate for the thousands of calls needed in real-time control of experiments and for use in integrated models, we will apply modern machine learning (ML) techniques to accelerate these models and interpolate their results. To generate the fast and accurate models for use in control level algorithms and integrated modeling we need to replace present models with high dimensional interpolation of their results. We will perform additional simulations across a broader parameter range for EAST and other tokamaks in different physics regimes (Alcator C-Mod, DIII-D, WEST, CFETR, ARC, ITER) and combine them into a larger database for training and testing of the ML models. Further testing of the control level models with experimental current profile data from EAST and C-Mod tokamaks will provide additional confirmation of the control level model before integration in a tokamak control system or integrated modeling suite. ML will be used to optimize the selection of training data consisting of RF current driven at different values of density profile, temperature profile, plasma current, and wavenumber. ML will also be used to facilitate classification of current drive from these input data. The output of this effort will be a validated classifier capable of determining the current drive profiles for HHFW CD and LHCD on a mille-second timescale. This will provide a breakthrough capability enabling real-time control of RF driven current profiles in experiments including ITER ICRF and use integrated modeling frameworks requiring thousands of current profile calculations in discharge simulations.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

A Wrapper to Use a Machine-Learning-Based Algorithm for Earthquake Monitoring

Seismology is one of the main sciences used to monitor volcanic activity worldwide. Fast, efficient, and accurate seismicity detectors are crucial to assess the activity level of a volcano in near–real time and to issue timely warnings. Traditional real–time seismic processing software uses phase onset pickers followed by a phase association algorithm to declare an event and estimate its location. The pickers typically do not identify whether the detected phase is a P or S arrival, which can have a negative impact on hypocentral location quality and complicates phase association. We implemented the deep–neural–network–based method PhaseNet to identify in real time P and S seismic waves on data from one– and three–component seismometers. We tuned the Earthworm binder_ew associator module to use the phase identification from PhaseNet to detect and locate the events, which we archive in a SeisComP3 database. We assessed the performance of the algorithm by comparing the results with existing catalogs built to monitor seismic and volcanic activity in Mayotte and the Lesser Antilles region. Our algorithm, which we refer to as PhaseWorm, showed promising results in both contexts and clearly outperformed the previous automatic method implemented in Mayotte. As a result, this innovative real–time processing system is now operational for seismicity monitoring in Mayotte and Martinique.

58 GEOSCIENCES↗

An augmented Lagrangian filter method

Here, we introduce a filter mechanism to enforce convergence for augmented Lagrangian methods for nonlinear programming. In contrast to traditional augmented Lagrangian methods, our approach does not require the use of forcing sequences that drive the first-order error to zero. Instead, we employ a filter to drive the optimality measures to zero. Our algorithm is flexible in the sense that it allows for equality-constrained quadratic programming steps to accelerate local convergence. We also include a feasibility restoration phase that allows fast detection of infeasible problems. We provide a convergence proof that shows that our algorithm converges to first-order stationary points. We provide preliminary numerical results that demonstrate the effectiveness of our proposed method.

97 MATHEMATICS AND COMPUTING↗

Online real-time learning of dynamical systems from noisy streaming data

Abstract Recent advancements in sensing and communication facilitate obtaining high-frequency real-time data from various physical systems like power networks, climate systems, biological networks, etc. However, since the data are recorded by physical sensors, it is natural that the obtained data is corrupted by measurement noise. In this paper, we present a novel algorithm for online real-time learning of dynamical systems from noisy time-series data, which employs the Robust Koopman operator framework to mitigate the effect of measurement noise. The proposed algorithm has three main advantages: (a) it allows for online real-time monitoring of a dynamical system; (b) it obtains a linear representation of the underlying dynamical system, thus enabling the user to use linear systems theory for analysis and control of the system; (c) it is computationally fast and less intensive than the popular extended dynamic mode decomposition (EDMD) algorithm. We illustrate the efficiency of the proposed algorithm by applying it to identify the Van der Pol oscillator, the chaotic attractor of the Henon map, the IEEE 68 bus system, and a ring network of Van der Pol oscillators.

97 MATHEMATICS AND COMPUTING↗