Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “distributed algorithm”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 631 records · Page 35

Enabling Efficient Sparse Computations using Linear Algebra Aware Compilers

This project developed the LAPIS compiler framework, built on the Multilevel Intermediate Representation (MLIR), to optimize sparse linear algebra operations and support performance portability across diverse architectures. The main innovation of LAPIS is the Kokkos dialect, which allows for lowering codes from a high productivity language to different architectures in an elegant way. The dialect also allows the conversion of lower-level MLIR code to C++ Kokkos code, facilitating the integration of scientific machine learning (SciML) models into applications. To extend LAPIS for distributed memory architectures, a new partition dialect was created to manage the distribution of sparse tensors and express communication patterns for sparse linear algebra operations. This dialect also supports the distributed execution of operators and includes algorithmic optimizations to minimize communication to improve performance. The project also demonstrates that MLIR can enable effective linear algebra-level optimizations, improving performance on different GPUs for both sparse and dense linear algebra kernels. Key applications of LAPIS include sparse linear algebra and graph kernels, TenSQL, a relational database management solution built on GraphBLAS, and the development of subgraph isomorphism and monomorphism kernels, showcasing performance portability. In summary, the LAPIS framework supports productivity, performance, portability, and distributed memory execution, while also enabling linear algebra-level optimizations that are challenging in traditional programming languages, with successful applications ranging from simple sparse linear algebra to complex graph kernels.

97 MATHEMATICS AND COMPUTING↗

SPHERES National Lab Facility

SPHERES is a facility of the ISS National Laboratory with three IVA nano-satellites designed and delivered by MIT to research estimation, control, and autonomy algorithms. Since Fall 2010, The SPHERES system is now operationally supported and managed by NASA Ames Research Center (ARC). A SPHERES Program Office was established and is located at NASA Ames Research Center. The SPHERES Program Office coordinates all SPHERES related research and STEM activities on-board the International Space Station (ISS), as well as, current and future payload development. By working aboard ISS under crew supervision, it provides a risk tolerant Test-bed Environment for Distributed Satellite Free-flying Control Algorithms. If anything goes wrong, reset and try again! NASA has made the capability available to other U.S. government agencies, schools, commercial companies and students to expand the pool of ideas for how to test and use these bowling ball-sized droids. For many of the researchers, SPHERES offers the only opportunity to do affordable on-orbit characterization of their technology in the microgravity environment. Future utilization of SPHERES as a facility will grow its capabilities as a platform for science, technology development, and education.

Free Flyer↗

NavP: Structured and Multithreaded Distributed Parallel Programming

We present Navigational Programming (NavP) -- a distributed parallel programming methodology based on the principles of migrating computations and multithreading. The four major steps of NavP are: (1) Distribute the data using the data communication pattern in a given algorithm; (2) Insert navigational commands for the computation to migrate and follow large-sized distributed data; (3) Cut the sequential migrating thread and construct a mobile pipeline; and (4) Loop back for refinement. NavP is significantly different from the current prevailing Message Passing (MP) approach. The advantages of NavP include: (1) NavP is structured distributed programming and it does not change the code structure of an original algorithm. This is in sharp contrast to MP as MP implementations in general do not resemble the original sequential code; (2) NavP implementations are always competitive with the best MPI implementations in terms of performance. Approaches such as DSM or HPF have failed to deliver satisfying performance as of today in contrast, even if they are relatively easy to use compared to MP; (3) NavP provides incremental parallelization, which is beyond the reach of MP; and (4) NavP is a unifying approach that allows us to exploit both fine- (multithreading on shared memory) and coarse- (pipelined tasks on distributed memory) grained parallelism. This is in contrast to the currently popular hybrid use of MP+OpenMP, which is known to be complex to use. We present experimental results that demonstrate the effectiveness of NavP.

navigational programming (NavP)↗

Gap-filling eddy covariance methane fluxes: Comparison of machine learning model predictions and uncertainties at FLUXNET-CH4 wetlands

Time series of methane fluxes measured by eddy-covariance require gap-filling to estimate annual emissions. Gap-filling methane fluxes is challenging because of high variability and complex responses to multiple drivers. To date, there is no widely established gap-filling standard for methane, with regards both to the best model algorithms and predictors. In this study, we address the need for standardization by synthesizing results of gap-filling methods applied at 17 wetland sites spanning boreal to tropical regions including all major wetlands classes and two rice paddies. We introduce new procedures for: 1) creating realistic artificial gap scenarios, 2) training and evaluating gap-filling models without overstating performance, and 3) predicting half-hourly methane fluxes and annual emissions with robust uncertainty estimates. We tested a conventional method (marginal distribution sampling) and four machine learning algorithms - penalized linear regression, artificial neural networks, random forests, and boosted decision trees - and four predictor sets, including temporal, meteorological, ecosystem carbon and energy flux, and soil predictors. We find that the conventional method can achieve similar median performance to the machine learning models but is worse than the best machine learning models and relatively insensitive to predictor choices. Of the machine learning models, decision tree algorithms performed the best in cross-validation experiments, even with a baseline predictor set, and artificial neural networks showed comparable performance when using all predictors. Soil temperature was frequently the most important predictor whilst water table depth was important at sites with substantial water table fluctuations, highlighting the value of data on soil conditions. Raw gap-filling uncertainties from the machine learning models were underestimated and we propose a method to calibrate uncertainties to observations. Finally, we gap-fill and provide summary evaluation metrics for all 81 sites in the FLUXNET-CH4 community dataset and publicly release the python code for model development, evaluation, and uncertainty estimation.

42 ENGINEERING↗

Intercomparison of Three Continuous Monitoring Systems on Operating Oil and Gas Sites

We compare continuous monitoring systems (CMS) from three different vendors on six operating oil and gas sites in the Appalachian Basin using several months of data. We highlight similarities and differences between the three CMS solutions when deployed in the field and compare their output to concurrent top-down aerial measurements and to site-level bottom-up inventories. Furthermore, we compare vendor-provided emission rate estimates to estimates from an open-source quantification algorithm applied to the raw CMS concentration data. This experimental setup allows us to separate the effect of the sensor platform (i.e., sensor type and arrangement) from the quantification algorithm. We find that 1) localization and quantification estimates rarely agree between the three CMS solutions on short time scales (i.e., 30 min), but temporally aggregated emission rate distributions are similar between solutions, 2) differences in emission rate distributions are generally driven by the quantification algorithm, rather than the sensor platform, 3) agreement between CMS and aerial rate estimates varies by CMS solution but is close to parity when CMS estimates are averaged across solutions, and 4) similar sites with similar bottom-up inventories do not necessarily have similar emission characteristics. These results have important implications for developing measurement-informed inventories and for incorporating CMS-inferred emission characteristics into emission mitigation efforts.

54 ENVIRONMENTAL SCIENCES↗

Learning to read aloud: A neural network approach using sparse distributed memory

An attempt to solve a problem of text-to-phoneme mapping is described which does not appear amenable to solution by use of standard algorithmic procedures. Experiments based on a model of distributed processing are also described. This model (sparse distributed memory (SDM)) can be used in an iterative supervised learning mode to solve the problem. Additional improvements aimed at obtaining better performance are suggested.

Joglekar, Umesh Dwarkanath↗

An extended numerical manifold method for unsaturated soil-water interaction analysis at micro-scale

To investigate unsaturated soil-water interaction at micro-scale, this work extends the numerical manifold method (NMM) by incorporating a soil-water coupling model considering specific capillary water distribution and capillary force calculation. The soil skeleton is constructed by a soil skeleton generation algorithm with random polygons. To more realistically capture the interaction between soil grains and capillary water, a capillary mechanics-based geometric algorithm is proposed to iteratively calculate the capillary water distribution. The capillary forces corresponding to the capillary water distribution are calculated based on the Young-Laplace equation. The proposed capillary water solving framework is first verified by reproducing the soil-water characteristic curve and the capillary water distribution of an ideal contact-disk model against analytical solutions. To further validate the ability of the capillary water solving framework to predict hydraulic behavior of the real soil, a laboratory test on the Toyoura sand is reproduced numerically. Then an ideal direct shear test is performed to further validate the two-way soil-water coupling procedure, in which a comparison between the numerical and analytical results regarding the shear strength and matric suction is presented. Finally, microscopic hydraulic and compression tests are conducted on two soil specimens with the same porosity and mean grain diameter but different uniformity coefficients. The results elucidate that the extended method is a potential tool to explore unsaturated soil behaviors at micro-scale.

58 GEOSCIENCES↗

Measuring the thermal and ionization state of the low- z IGM using likelihood free inference

ABSTRACT We present a new approach to measure the power-law temperature density relationship $T=T_0 (\rho/ \bar{\rho })^{\gamma -1}$ and the UV background photoionization rate $\Gamma _{{{{\rm H\, {\small I}}}}{}}$ of the intergalactic medium (IGM) based on the Voigt profile decomposition of the Ly α forest into a set of discrete absorption lines with Doppler parameter b and the neutral hydrogen column density $N_{\rm H\, {\small I}}$. Previous work demonstrated that the shape of the $b-N_{{{{\rm H\, {\small I}}}}{}}$ distribution is sensitive to the IGM thermal parameters T0 and γ, whereas our new inference algorithm also takes into account the normalization of the distribution, i.e. the line-density dN/dz, and we demonstrate that precise constraints can also be obtained on $\Gamma _{{{{\rm H\, {\small I}}}}{}}$. We use density-estimation likelihood-free inference (DELFI) to emulate the dependence of the $b-N_{{{{\rm H\, {\small I}}}}{}}$ distribution on IGM parameters trained on an ensemble of 624 nyx hydrodynamical simulations at z = 0.1, which we combine with a Gaussian process emulator of the normalization. To demonstrate the efficacy of this approach, we generate hundreds of realizations of realistic mock HST/COS data sets, each comprising 34 quasar sightlines, and forward model the noise and resolution to match the real data. We use this large ensemble of mocks to extensively test our inference and empirically demonstrate that our posterior distributions are robust. Our analysis shows that by applying our new approach to existing Ly α forest spectra at z ≃ 0.1, one can measure the thermal and ionization state of the IGM with very high precision ($\sigma _{\log T_0} \sim 0.08$ dex, σγ ∼ 0.06, and $\sigma _{\log \Gamma _{{{{\rm H\, {\small I}}}}{}}} \sim 0.07$ dex).

79 ASTRONOMY AND ASTROPHYSICS↗

Quantum Markov chain Monte Carlo with digital dissipative dynamics on quantum computers

Modeling the dynamics of a quantum system connected to the environment is critical for advancing our understanding of complex quantum processes, as most quantum processes in nature are affected by an environment. Modeling a macroscopic environment on a quantum simulator may be achieved by coupling independent ancilla qubits that facilitate energy exchange in an appropriate manner with the system and mimic an environment. This approach requires a large, and possibly exponential number of ancillary degrees of freedom which is impractical. In contrast, we develop a digital quantum algorithm that simulates interaction with an environment using a small number of ancilla qubits. By combining periodic modulation of the ancilla energies, or spectral combing, with periodic reset operations, we are able to mimic interaction with a large environment and generate thermal states of interacting many-body systems. We evaluate the algorithm by simulating preparation of thermal states of the transverse Ising model. Our algorithm can also be viewed as a quantum Markov chain Monte Carlo process that allows sampling of the Gibbs distribution of a multivariate model. To illustrate this we evaluate the accuracy of sampling Gibbs distributions of simple probabilistic graphical models using the algorithm.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Evaluation of four ground-based retrievals of cloud droplet number concentration in marine stratocumulus with aircraft in situ measurements

Abstract. Cloud droplet number concentration (Nd) is crucial for understanding aerosol–cloud interactions (ACI) and associated radiative effects. We present evaluations of four ground-based Nd retrievals based on comprehensive datasets from the Atmospheric Radiation Measurement (ARM) Aerosol and Cloud Experiments in the Eastern North Atlantic (ACE-ENA) field campaign. The Nd retrieval methods use ARM ENA observatory ground-based remote sensing observations from a micropulse lidar, Raman lidar, cloud radar, and the ARM NDROP (Droplet Number Concentration) value-added product (VAP), all of which also retrieve cloud effective radius (re). The retrievals are compared against aircraft measurements from the fast cloud droplet probe (FCDP) and the cloud and aerosol spectrometer (CAS) obtained from low-level marine boundary layer clouds on 12 flight days during summer and winter seasons. Additionally, the in situ measurements are used to validate the assumptions and characterizations used in the retrieval algorithms. Statistical comparisons of the probability distribution function (PDF) of the Nd and cloud re retrievals with aircraft measurements demonstrate that these retrievals align well with in situ measurements for overcast clouds, but they may substantially differ for broken clouds or clouds with low liquid water path (LWP). The retrievals are applied to 4 years of ground-based remote sensing measurements of overcast marine boundary layer clouds at the ARM ENA observatory to find that Nd (re) values exhibit seasonal variations, with higher (lower) values during the summer season and lower (higher) values during the winter season. The ensemble of various retrievals using different measurements and retrieval algorithms such as those in this paper can help to quantify Nd retrieval uncertainties and identify reliable Nd retrieval scenarios. Of the retrieval methods, we recommend using the micropulse lidar-based method. This method has good agreement with in situ measurements, less sensitivity to issues arising from precipitation and low cloud LWP and/or optical depth, and broad applicability by functioning for both daytime and nighttime conditions.

54 ENVIRONMENTAL SCIENCES↗

GOCI Yonsei Aerosol Retrieval (YAER) Algorithm and Validation During the DRAGON-NE Asia 2012 Campaign

The Geostationary Ocean Color Imager (GOCI) onboard the Communication, Ocean, and Meteorological Satellite (COMS) is the first multi-channel ocean color imager in geostationary orbit. Hourly GOCI top-of-atmosphere radiance has been available for the retrieval of aerosol optical properties over East Asia since March 2011. This study presents improvements made to the GOCI Yonsei Aerosol Retrieval (YAER) algorithm together with validation results during the Distributed Regional Aerosol Gridded Observation Networks - Northeast Asia 2012 campaign (DRAGONNE Asia 2012 campaign). The evaluation during the spring season over East Asia is important because of high aerosol concentrations and diverse types of Asian dust and haze. Optical properties of aerosol are retrieved from the GOCI YAER algorithm including aerosol optical depth (AOD) at 550 nm, fine-mode fraction (FMF) at 550 nm, single-scattering albedo (SSA) at 440 nm, Angstrom exponent (AE) between 440 and 860 nm, and aerosol type. The aerosol models are created based on a global analysis of the Aerosol Robotic Networks (AERONET) inversion data, and covers a broad range of size distribution and absorptivity, including nonspherical dust properties. The Cox-Munk ocean bidirectional reflectance distribution function (BRDF) model is used over ocean, and an improved minimum reflectance technique is used over land. Because turbid water is persistent over the Yellow Sea, the land algorithm is used for such cases. The aerosol products are evaluated against AERONET observations and MODIS Collection 6 aerosol products retrieved from Dark Target (DT) and Deep Blue (DB) algorithms during the DRAGON-NE Asia 2012 campaign conducted from March to May 2012. Comparison of AOD from GOCI and AERONET resulted in a Pearson correlation coefficient of 0.881 and a linear regression equation with GOCI AOD = 1.083 x AERONET AOD - 0.042. The correlation between GOCI and MODIS AODs is higher over ocean than land. GOCI AOD shows better agreement with MODIS DB than MODIS DT. The other GOCI YAER products (AE, FMF, and SSA) show lower correlation with AERONET than AOD, but still show some skills for qualitative use.

Choi, Myungje↗

Evaluation of Four Ground-based Retrievals of Cloud Droplet Number Concentration in Marine Stratocumulus with Aircraft In Situ Measurements

Cloud droplet number concentration (N d ) is crucial for understanding aerosol-cloud interactions (ACI) and associated radiative effects. We present evaluations of four ground-based N d retrievals based on comprehensive datasets from the Atmospheric Radiation Measurements (ARM) Aerosol and Cloud Experiments in the Eastern North Atlantic (ACE-ENA) field campaign. The N d retrieval methods use ARM ENA observatory ground-based remote sensing observations from a Micropulse lidar, Raman lidar, cloud radar, and the ARM NDROP Value-added Product (VAP), all of which also retrieve cloud effective radius (r e ). The retrievals are compared against aircraft measurements from the Fast-Cloud Droplet Probe (FCDP) and the Cloud and Aerosol Spectrometer (CAS) obtained from low-level marine boundary layer clouds on 12 flight days during summer and winter seasons. Additionally, the in situ measurements are used to validate the assumptions and characterizations used in the retrieval algorithms. Statistical comparisons of the probability distribution function (PDF) of the N d and cloud r e retrievals with aircraft measurements demonstrate that these retrievals align well with in situ measurements for overcast clouds, but they may substantially differ for broken clouds or clouds with low liquid water path (LWP). The retrievals are applied to four years of ground-based remote sensing measurements of overcast marine boundary layer clouds at the ARM ENA observatory to find that N d (r e ) values exhibit seasonal variations, with higher (lower) values during the summer season and lower (higher) values during the winter season. The ensemble of various retrievals using different measurements and retrieval algorithms such as those in this paper can help to quantify N d retrieval uncertainties and identify reliable N d retrieval scenarios. Of the retrieval methods, we recommend using the using the Micropulse lidar-based method given its good agreement with in situ measurements, it has less sensitivity to issues arising from precipitation and low cloud LWP/optical depth, and it has broad applicability by functioning for both day and nighttime conditions.

54 ENVIRONMENTAL SCIENCES↗

The OMPS Limb Profiler Instrument: An Alternative Data Analysis and Retrieval Algorithm

The upcoming Ozone Mapper and Profiler Suite (OMPS), which will be launched on the NPOESS Preparatory Project (NPP) platform in early 2011, will continue monitoring the global distribution of the Earth's middle atmosphere ozone and aerosol. OMPS is composed of three instruments, namely the Total Column Mapper (heritage: TOMS, OMI), the Nadir Profiler (heritage: SBUV) and the Limb Profiler (heritage: SOLSE/LORE, OSIRIS, SCIAMACHY, SAGE III). The ultimate goal of the mission is to better understand and quantify the rate of stratospheric ozone recovery. The focus of the paper will be on the Limb Profiler (LP) instrument. The LP instrument will measure the Earth fs limb radiance (which is due to the scattering of solar photons by air molecules, aerosol and Earth surface) in the ultra-violet (UV), visible and near infrared, from 285 to 1000 nm. The LP simultaneously images the whole vertical extent of the Earth's limb through three vertical slits, each covering a vertical tangent height range of 100 km and each horizontally spaced by 250 km in the cross-track direction. The focal plane of the LP spectrometer is a two ]dimensional CCD array comprised of 340 x 740 pixels. Several data analysis tools are presently being constructed and tested to retrieve ozone and aerosol vertical distribution from limb radiance measurements. The primary NASA algorithm is based on earlier algorithms developed for the SOLSE/LORE and SAGE III limb scatter missions. The paper will describe an alternative algorithm which will retrieve ozone density and aerosol extinction directly from radiance data collected on individual CCD pixels. This alternative method uses an optimal estimation approach to retrieve ozone and aerosol in the 10-60 km range from the information contained within an ensemble of about 50000 down-linked pixels. Tangent height registration is performed using the Rayleigh Scattering Attitude Sensor (RSAS) technique applied to columns of pixels in the 340-360 nm range. Cloud height is determined by analyzing the radiance first derivative along pixel columns at longer wavelengths. Wavelength registration is performed using rows of pixels and identifying Fraunhofer solar lines within the measured spectra. Special attention is given to stray-light decontamination and modeling of the measured finite spectral/spatial line shape functions.

Rault, Didier F.↗

A Hybrid Optimization and Deep Learning Algorithm for Cyber-Resilient DER Control

With the proliferation of distributed energy resources (DERs) in the distribution grid, it is a challenge to effectively control a large number of DERs resilient to the communication and security disruptions, as well as to provide the online grid services, such as voltage regulation and virtual power plant (VPP) dispatch. To this end, a hybrid feedback-based optimization algorithm along with deep learning forecasting technique is proposed to specifically address the cyber-related issues. The online decentralized feedback-based DER optimization control requires timely, accurate voltage measurement from the grid. However, in practice such information may not be received by the control center or even be corrupted. Therefore, the long short-term memory (LSTM) deep learning algorithm is employed to forecast delayed/missed/attacked messages with high accuracy. The IEEE 37-node feeder with high penetration of PV systems is used to validate the efficiency of the proposed hybrid algorithm. The results show that 1) the LSTM-forecasted lost voltage can effectively improve the performance of the DER control algorithm in the practical cyber-physical architecture; and 2) the LSTM forecasting strategy outperforms other strategies of using previous message and skipping dual parameter update.

cyber-resilient algorithm↗

GA-Based Voltage Optimization of Distribution Feeder with High-Penetration of DERs Using Megawatt-Scale Units

In this paper, genetic algorithm (GA)-based voltage optimization of a modified IEEE-34 node distribution feeder with high penetration of distributed energy resources (DERs) is proposed using two megawatt-scale reactive power sources. Traditional voltage support units present in distribution grids are not suitable for DER-rich feeders, while voltage support using small-scale DERs present in the feeder requires considerable communication effort to reach a global solution. In this work, two megawatt-scale units are placed to improve the voltage profile across the IEEE 34-node feeder, which has been modified to include several PV units and an energy storage unit. The megawatt-scale units are optimized using GA for fast and accurate operation. The performance of the proposed scheme is verified using simulation results with a multi-platform setup where the modified IEEE-34 node feeder is modeled in OpenDSS while the GA optimization scheme is programmed in MATLAB.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Evaluation of a pulse control law for flexible spacecraft

The following analytical and experimental studies were conducted: (1) A simple algorithm was developed to suppress the structural vibrations of 3-dimensional distributed parameter systems, subjected to interface motion and/or directly applied forces. The algorithm is designed to cope with structural oscillations superposed on top of rigid-body motion: a situation identical to that encountered by the SCOLE components. A significant feature of the method is that only local measurements of the structural displacements and velocities relative to the moving frame of reference are needed. (2) A numerical simulation study was conducted on a simple linear finite element model of a cantilevered plate which was subjected to test excitations consisting of impulsive base motion and of nonstationary wide-band random excitation applied at its root. In each situation, the aim was to suppress the vibrations of the plate relative to the moving base. (3) A small mechanical model resembling an aircraft wing was designed and fabricated to investigate the control algorithm under realistic laboratory conditions.

Source record↗

Nearly optimal state preparation for quantum simulations of lattice gauge theories

Here, we present several improvements to the recently developed ground-state preparation algorithm based on the quantum eigenvalue transformation for unitary matrices (QETU), apply this algorithm to a lattice formulation of U(1) gauge theory in (2+1) dimensions, as well as propose an alternative application of QETU, a highly efficient preparation of Gaussian distributions. The QETU technique was originally proposed as an algorithm for nearly optimal ground-state preparation and ground-state energy estimation on early fault-tolerant devices. It uses the time-evolution input model, which can potentially overcome the large overall prefactor in the asymptotic gate cost arising in similar algorithms based on the Hamiltonian input model. We present modifications to the original QETU algorithm that significantly reduce the cost for the cases of both exact and Trotterized implementation of the time evolution circuit. We use QETU to prepare the ground state of a U(1) lattice gauge theory in two spatial dimensions, explore the dependence of computational resources on the desired precision and system parameters, and discuss the applicability of our results to general lattice gauge theories. We also demonstrate how the QETU technique can be utilized for preparing Gaussian distributions and wave packets in a way which outperforms existing algorithms for as little as n q ≳ 2–5 qubits.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Multi-Level Optimal Power Flow Solver in Large Distribution Networks: Preprint

Solving optimal power flow (OPF) problem for large distribution networks incurs high computational complexity. We consider a large multi-phase distribution networks of tree topology with deep penetration of active devices. We divide the network into collaborating areas featuring subtree topology and subareas featuring subsubtree topology. We design a multi-level implementation of the primal-dual gradient algorithm for solving the voltage regulation OPF problems while preserving nodal voltage information and topological information within areas and subareas. Numerical results on a 4,521-node system verifies that the proposed algorithm can significantly improve computational speed without compromising any optimality.

41 EE - Solar Energy Technologies Office (EE-4S)↗