Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “approximate computing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Data-driven wind turbine wake modeling via probabilistic machine learning

Wind farm design primarily depends on the variability of the wind turbine wake flows to the atmospheric wind conditions and the interaction between wakes. Physics-based models that capture the wake flow field with high-fidelity are computationally very expensive to perform layout optimization of wind farms, and, thus, data-driven reduced-order models can represent an efficient alternative for simulating wind farms. In this work, we use real-world light detection and ranging (LiDAR) measurements of wind-turbine wakes to construct predictive surrogate models using machine learning. Specifically, we first demonstrate the use of deep autoencoders to find a low-dimensional latent space that gives a computationally tractable approximation of the wake LiDAR measurements. Then, we learn the mapping between the parameter space and the (latent space) wake flow fields using a deep neural network. Additionally, we also demonstrate the use of a probabilistic machine learning technique, namely, Gaussian process modeling, to learn the parameter-space-latent-space mapping in addition to the epistemic and aleatoric uncertainty in the data. Finally, to cope with training large datasets, we demonstrate the use of variational Gaussian process models that provide a tractable alternative to the conventional Gaussian process models for large datasets. Furthermore, we introduce the use of active learning to adaptively build and improve a conventional Gaussian process model predictive capability. Overall, we find that our approach provides accurate approximations of the wind-turbine wake flow field that can be queried at an orders-of-magnitude cheaper cost than those generated with high-fidelity physics-based simulations.

Deep neural networks↗

Online multimedia retrieval on CPU–GPU platforms with adaptive work partition

Nearest neighbors search is a core operation found in several online multimedia services. These services have to handle very large databases, while, at the same time, they must minimize the query response times observed by users. This is specially complex because those services deal with fluctuating query workloads (rates). Consequently, they must adapt at run-time to minimize the response times as the load varies. In this paper, we address the aforementioned challenges with a distributed memory parallelization of the product quantization nearest neighbor search, also known as IVFADC, for hybrid CPU–GPU machines. Overall, our parallel IVFADC implements an out-of-GPU memory execution scheme to use the GPU for databases in which the index does not fit in its memory, which is crucial for searching in very large databases. The careful use of CPU and GPU with work stealing led to an average response time reduction of 2.4 as compared to using the GPU only. Also, our approach to adapt the system to fluctuating loads, called Dynamic Query Processing Policy (DQPP), attained a response time reduction of up to 5 vs. the best static (BS) policy for moderate loads. The system has attained high query processing rates and near-linear scalability in all experiments. We have evaluated our system on a machine with up to 256 NVIDIA V100 GPUs processing a database of 256 billion SIFT features vectors.

97 MATHEMATICS AND COMPUTING↗

A Hybrid Method for Tensor Decompositions that Leverages Stochastic and Deterministic Optimization

In this paper, we propose a hybrid method that uses stochastic and deterministic search to compute the maximum likelihood estimator of a low-rank count tensor with Poisson loss via state-of-theart local methods. Our approach is inspired by Simulated Annealing for global optimization and allows for fine-grain parameter tuning as well as adaptive updates to algorithm parameters. We present numerical results that indicate our hybrid approach can compute better approximations to the maximum likelihood estimator with less computation than the state-of-the-art methods by themselves.

97 MATHEMATICS AND COMPUTING↗

Random insights into the complexity of two-dimensional tensor network calculations

Projected entangled pair states (PEPS) offer memory-efficient representations of some quantum many-body states that obey an entanglement area law and are the basis for classical simulations of ground states in two-dimensional (2d) condensed matter systems. However, rigorous results show that exactly computing observables from a 2d PEPS state is generically a computationally hard problem. Yet approximation schemes for computing properties of 2d PEPS are regularly used, and empirically seen to succeed, for a large subclass of (“not too entangled”) condensed matter ground states. Adopting the philosophy of random matrix theory, in this work, we analyze the complexity of approximately contracting a 2d random PEPS by exploiting an analytic mapping to an effective replicated statistical mechanics model that permits a controlled analysis at a large bond dimension. Through this statistical-mechanics lens, we argue that (i) although approximately sampling wave-function amplitudes of random PEPS faces a computational-complexity phase transition above a critical bond dimension, and (ii) one can generically efficiently estimate the norm and correlation functions for any finite bond dimension. Furthermore, these results are supported numerically for various bond-dimension regimes. It is an important open question whether the above results for random PEPS apply more generally also to PEPS representing physically relevant ground states.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Recognizability of Demographically Altered Computerized Facial Approximations in an Automated Facial Recognition Context for Potential Application in Unidentified Persons Data Repositories

This study examined the recognizability of demographically altered facial approximations for potential utility in unidentified persons tracking systems. Five computer-generated approximations were generated for each of 26 African male participants using the following demographic parameters: (i) African male (true demographics), (ii) African female, (iii) Caucasian male, (iv) Asian male, and (v) Hispanic male. Overall, 62% of the true demographic facial approximations for the 26 African male participants examined were matched to a corresponding life photo within the top 50 images of a candidate list generated from an automated blind search of an optimally standardized gallery of 6159 photographs. When the African male participants were processed as African females, the identification rate was 50%. In contrast, less congruent identification rates were observed when the African male participants were processed as Caucasian (42%), Asian (35%), and Hispanic (27%) males. The observed results suggest that approximations generated using the opposite sex may be operationally informative if sex is unknown. The performance of approximations generated using alternative ancestry assignments, however, was less congruent with the performance of the true demographic approximation (African male) and may not yield as operationally constructive data as sex-altered approximations.

59 BASIC BIOLOGICAL SCIENCES↗

A Stochastic Intracellular Model of Anthrax Infection With Spore Germination Heterogeneity

We present a stochastic mathematical model of the intracellular infection dynamics of Bacillus anthracis in macrophages. Following inhalation of B. anthracis spores, these are ingested by alveolar phagocytes. Ingested spores then begin to germinate and divide intracellularly. This can lead to the eventual death of the host cell and the extracellular release of bacterial progeny. Some macrophages successfully eliminate the intracellular bacteria and will recover. Here, a stochastic birth-and-death process with catastrophe is proposed, which includes the mechanism of spore germination and maturation of B. anthracis . The resulting model is used to explore the potential for heterogeneity in the spore germination rate, with the consideration of two extreme cases for the rate distribution: continuous Gaussian and discrete Bernoulli. We make use of approximate Bayesian computation to calibrate our model using experimental measurements from in vitro infection of murine peritoneal macrophages with spores of the Sterne 34F2 strain of B. anthracis . The calibrated stochastic model allows us to compute the probability of rupture, mean time to rupture, and rupture size distribution, of a macrophage that has been infected with one spore. We also obtain the mean spore and bacterial loads over time for a population of cells, each assumed to be initially infected with a single spore. Our results support the existence of significant heterogeneity in the germination rate, with a subset of spores expected to germinate much later than the majority. Furthermore, in agreement with experimental evidence, our results suggest that most of the spores taken up by macrophages are likely to be eliminated by the host cell, but a few germinated spores may survive phagocytosis and lead to the death of the infected cell. Finally, we discuss how this stochastic modelling approach, together with dose-response data, allows us to quantify and predict individual infection risk following exposure.

59 BASIC BIOLOGICAL SCIENCES↗

Inferring pesticide toxicity to honey bees from a field‐based feeding study using a colony model and Bayesian inference

Abstract Honey bees are crucial pollinators for agricultural crops but are threatened by a multitude of stressors including exposure to pesticides. Linking our understanding of how pesticides affect individual bees to colony‐level responses is challenging because colonies show emergent properties based on complex internal processes and interactions among individual bees. Agent‐based models that simulate honey bee colony dynamics may be a tool for scaling between individual and colony effects of a pesticide. The U.S. Environmental Protection Agency (USEPA) and U.S. Department of Agriculture (USDA) are developing the VarroaPop + Pesticide model, which simulates the dynamics of honey bee colonies and how they respond to multiple stressors, including weather, Varroa mites, and pesticides. To evaluate this model, we used Approximate Bayesian Computation to fit field data from an empirical study where honey bee colonies were fed the insecticide clothianidin. This allowed us to reproduce colony feeding study data by simulating colony demography and mortality from ingestion of contaminated food. We found that VarroaPop + Pesticide was able to fit general trends in colony population size and structure and reproduce colony declines from increasing clothianidin exposure. The model underestimated adverse effects at low exposure (36 µg/kg), however, and overestimated recovery at the highest exposure level (140 µg/kg), for the adult and pupa endpoints, suggesting that mechanisms besides oral toxicity‐induced mortality may have played a role in colony declines. The VarroaPop + Pesticide model estimates an adult oral LD 50 of 18.9 ng/bee (95% CI 10.1–32.6) based on the simulated feeding study data, which falls just above the 95% confidence intervals of values observed in laboratory toxicology studies on individual bees. Overall, our results demonstrate a novel method for analyzing colony‐level data on pesticide effects on bees and making inferences on pesticide toxicity to individual bees.

59 BASIC BIOLOGICAL SCIENCES↗

QUBO formulations for training machine learning models

Abstract Training machine learning models on classical computers is usually a time and compute intensive process. With Moore’s law nearing its inevitable end and an ever-increasing demand for large-scale data analysis using machine learning, we must leverage non-conventional computing paradigms like quantum computing to train machine learning models efficiently. Adiabatic quantum computers can approximately solve NP-hard problems, such as the quadratic unconstrained binary optimization (QUBO), faster than classical computers. Since many machine learning problems are also NP-hard, we believe adiabatic quantum computers might be instrumental in training machine learning models efficiently in the post Moore’s law era. In order to solve problems on adiabatic quantum computers, they must be formulated as QUBO problems, which is very challenging. In this paper, we formulate the training problems of three machine learning models—linear regression, support vector machine (SVM) and balanced k-means clustering—as QUBO problems, making them conducive to be trained on adiabatic quantum computers. We also analyze the computational complexities of our formulations and compare them to corresponding state-of-the-art classical approaches. We show that the time and space complexities of our formulations are better (in case of SVM and balanced k-means clustering) or equivalent (in case of linear regression) to their classical counterparts.

97 MATHEMATICS AND COMPUTING↗

The Thermophysical Properties of TcO2

Technetium-99 is a highly radioactive isotope with a long half-life that is common in nuclear waste. It volatizes at a low temperature, which poses a significant challenge to the clean-up and containment processes. Due to difficulties in purifying technetium compounds, their thermophysical properties have not been measured or calculated. Here, first principle methods are used along with the quasi quasi-harmonic harmonic approximation to compute the Debye temperature, volumetric thermal expansion coefficient, bulk modulus, and heat capacity of rutile TcO2 for temperatures ranging from 0 to 1500 K and applied pressures ranging from 0 to 255 GPa. The computed atomic structures agree well with the results from diffraction measurements. The computed thermophysical properties are in the neighborhood of other rutile metal oxides and, in particular, are within approximately 10–13% of rutile ReO2, which is frequently used as a substitute for TcO2 in experimental studies.

Zhong, Hong↗

TWST Cloud Properties (COD, droplet effective radius, and thermodynamic phase). SAIL, AMF2, 2022-2023.

The dataset consists of calibrated zenith shortwave spectral radiances (440-1700 nm) and cloud optical depth (COD) retrievals approximately every second. Retrievals of cloud droplet effective radius and thermodynamic phase are currently done offline, but some of these retrievals are included in the dataset. Effective radius and phase are included for days with periods of sufficiently high COD. Effective radius and phase retrievals in the data set were computed at approximately 5-10 second intervals to reduce computation time. The retrievals can be performed at the full sample rate. The field of view is 0.5 degree.

54 ENVIRONMENTAL SCIENCES↗

Thorium and Rare Earth Monoxides and Related Phases

Thorium was a part of energy infrastructure in the 19th century due to the refractory and electronic properties of its dioxide. It will be a part of future energy infrastructure as the most abundant energy reserve based on nuclear fission. This paper discusses the solid-state chemistry of the monoxides and related rocksalt phases of thorium and the rare earths, both at atmospheric and at high pressure. The existence of solid thorium monoxide was first suggested more than 100 years ago; however, it was never obtained in bulk and has been studied mostly theoretically. Monoxides of lanthanides from Eu to Ho are ferromagnetic semiconductors sought for spintronics and were studied in thin films. La to Sm metallic monoxides were synthesized in bulk at pressures below 5 GPa. Recently, ThO formation in thin films has been reported and the stability of bulk ThO at high pressure was theoretically predicted based on first principles computations at 0 K. New ab initio computations were performed accounting for temperature effects up to 1000 K using lattice dynamics in the quasi-harmonic approximation. New computational results confirm the stabilization of pure ThO above 30 GPa and suggest the possibility of high-pressure synthesis of (Th,Nd)O at 1000 K and 5 GPa.

36 MATERIALS SCIENCE↗

Conditional Point Sampling: A Monte Carlo Method for Radiation Transport in Stochastic Media.

Current methods for stochastic media transport are either computationally expensive or, by nature, approximate. Moreover, none of the well-developed, benchmarked approximate methods can compute the variance caused by the stochastic mixing, a quantity especially important to safety calculations. Therefore, we derive and apply a new conditional probability function (CPF) for use in the recently developed stochastic media transport algorithm Conditional Point Sampling (CoPS), which 1) leverages the full intra-particle memory of CoPS to yield errorless computation of stochastic media outputs in 1D, binary, Markovian-mixed media, and 2) leverages the full inter-particle memory of CoPS and the recently developed Embedded Variance Deconvolution method to yield computation of the variance in transport outputs caused by stochastic material mixing. Numerical results demonstrate errorless stochastic media transport as compared to reference benchmark solutions with the new CPF for this class of stochastic mixing as well as the ability to compute the variance caused by the stochastic mixing via CoPS. Using previously derived, non-errorless CPFs, CoPS is further found to be more accurate than the atomic mix approximation, Chord Length Sampling (CLS), and most of memory-enhanced versions of CLS surveyed. In addition, we study the compounding behavior of CPF error as a function of cohort size (where a cohort is a group of histories that share intra-particle memory) and recommend that small cohorts be used when computing the variance in transport outputs caused by stochastic mixing.

61 RADIATION PROTECTION AND DOSIMETRY↗

Aerodynamic Sensitivity of a Novel Data-Driven Airfoil Shape Representation Framework

We explore the aerodynamic implications of a novel data-driven separable shape tensor framework used to represent discrete airfoil shapes. In this study, we construct a data-driven parameter space defined by separable shape tensors and informed by tens of thousands of distinct airfoils. We use this design space to generate new airfoil designs to study parametric sensitivities with respect to various aerodynamic responses. We use a HAM2D RANS solver to approximate the lift, drag, and moment coefficients for the generated airfoils at two different angles-of-attack. We analyze the robustness and sensitivities of using the separable shape tensor design space by examining the coverage of the aerodynamic response space, uncovering low-dimensional polynomial ridge approximations, and computing various sensitivity metrics. The results show that the data-driven design space produce significant variation in target aerodynamic quantities and facilitate highly accurate approximations (R^2 > 0.96) of one- and two-dimensional structures in each aerodynamic response. This further reduces the effective dimension to enable simplified design and optimization tasks.

aerodynamics↗

Towards replacing physical testing of granular materials with a Topology-based Model

In the study of packed granular materials, the performance of a sample (e.g., the detonation of a high-energy explosive) often correlates to measurements of a fluid flowing through it. The “effective surface area,” the surface area accessible to the airflow, is typically measured using a permeametry apparatus that relates the flow conductance to the permeable surface area via the Carman-Kozeny equation. This equation allows calculating the flow rate of a fluid flowing through the granules packed in the sample for a given pressure drop. However, Carman-Kozeny makes inherent assumptions about tunnel shapes and flow paths that may not accurately hold in situations where the particles possess a wide distribution in shapes, sizes, and aspect ratios, as is true with many powdered systems of technological and commercial interest. To address this challenge, we replicate these measurements virtually on micro-CT images of the powdered material, introducing a new Pore Network Model based on the skeleton of the Morse-Smale complex. Pores are identified as basins of the complex, their incidence encodes adjacency, and the conductivity of the capillary between them is computed from the cross-section at their interface. We build and solve a resistive network to compute an approximate laminar fluid flow through the pore structure. Here, we provide two means of estimating flow-permeable surface area: (i) by direct computation of conductivity, and (ii) by identifying dead-ends in the flow coupled with isosurface extraction and the application of the Carman-Kozeny equation, with the aim of establishing consistency over a range of particle shapes, sizes, porosity levels, and void distribution patterns.

36 MATERIALS SCIENCE↗

Robust Implicit Adaptive Low Rank Time-Stepping Methods for Matrix Differential Equations

In this work, we develop implicit rank-adaptive schemes for time-dependent matrix differential equations. The dynamic low rank approximation (DLRA) is a well-known technique to capture the dynamic low rank structure based on Dirac–Frenkel time-dependent variational principle. In recent years, it has attracted a lot of attention due to its wide applicability. Our schemes are inspired by the three-step procedure used in the rank adaptive version of the unconventional robust integrator (the so called BUG integrator) (Ceruti et al. in BIT Numer Math 62(4):1149–1174, 2022) for DLRA. First, a prediction (basis update) step is made computing the approximate column and row spaces at the next time level. Second, a Galerkin evolution step is invoked using an implicit solves for the small core matrix. Finally, a truncation is made according to a prescribed error threshold. Since the DLRA is evolving the differential equation projected on to the tangent space of the low rank manifold, the error estimate of the BUG integrator contains the tangent projection (modeling) error which cannot be easily controlled by mesh refinement. This can cause convergence issue for equations with cross terms. To address this issue, we propose a simple modification, consisting of merging the row and column spaces from the explicit step truncation method together with the BUG spaces in the prediction step. In addition, we propose an adaptive strategy where the BUG spaces are only computed if the residual for the solution obtained from the prediction space by explicit step truncation method, is too large. Here, we prove stability and estimate the local truncation error of the schemes under assumptions. We benchmark the schemes in several tests, such as anisotropic diffusion, solid body rotation and the combination of the two, to show robust convergence properties.

97 MATHEMATICS AND COMPUTING↗

Effect of parallel flow on resonant layer responses in high beta plasmas

Abstract Resonant layers in a tokamak respond to non-axisymmetric magnetic perturbations by amplifying the mode amplitude and balancing the plasma rotation through magnetic reconnection and force balance, respectively. This resonant response can be characterized by local layer parameters and especially by a single quantity in the linear regime, the so-called inner-layer Δ. The computation of Δ under two-fluid drift-MHD formalism has been progressed by reducing the order of the system in the phase space, where the shielding current is approximated as being only carried by electrons, a posteriori . In this study, we relax the approximation and compute Δ accounted for by the parallel flow associated with the ion shielding current. The posteriori is numerically verified in great agreement with the original SLAYER developed in a previous paper (J.-K. Park 2022 Phys. Plasmas 29 072506). Extending the resonant layer response theory to high β plasmas, our research findings answer two important questions: how the parallel flow influences the resonant layer response and why the parallel flow effect appears in high β plasmas. The complicated plasma compression in high β regime allows the parallel flow response to give rise to the ion shielding current, which not only shifts the zero-crossing condition of the ExB flow but also enhances the field penetration threshold. Technically, the Riccati matrix transformation method is adapted to handle the numerical stiffness due to the increased order of the system. The high fidelity of this numerical method makes use of further extension of the model to higher-order systems to take other physical phenomena into account. This work is envisaged to predict the resonant layer response under high β fusion reactor conditions.

Lee, Yeongsun (ORCID:000000034474416X)↗

Direct Nonlinear Approximation for Security Region Boundary of Integrated Energy Systems: A Polynomial Chaos Expansion Solution

The strong interdependence of electricity, gas, and heating systems can facilitate fault propagation within integrated energy systems (IESs), posing significant challenges to secure operation. This paper proposes a polynomial chaos expansion (PCE)-based approximation method to accurately characterize the IES security region boundary (IES–SRB). By integrating the Karush-Kuhn-Tucker conditions with PCE theory, the IES-SRB approximation problem is reformulated as a set of nonlinear equations concerning the approximation coefficients. Using the Galerkin projection method, these equations are further transformed into a system of projection equations that govern the polynomial approximation coefficients in the IES-SRB approximation. To reduce computational complexity while maintaining high approximation accuracy, a piecewise polynomial approximation method is proposed. Numerical studies on the E39-G20-H6 and E118-G96-H52 IES test systems demonstrate that the proposed method can accurately and effectively construct IES security regions.

Wu, Chenghao [Northeast Electric Power University]↗