Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “hardware algorithms”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 271 records · Page 15

Constrained quantum optimization for extractive summarization on a trapped-ion quantum computer

Abstract Realizing the potential of near-term quantum computers to solve industry-relevant constrained-optimization problems is a promising path to quantum advantage. In this work, we consider the extractive summarization constrained-optimization problem and demonstrate the largest-to-date execution of a quantum optimization algorithm that natively preserves constraints on quantum hardware. We report results with the Quantum Alternating Operator Ansatz algorithm with a Hamming-weight-preserving XY mixer (XY-QAOA) on trapped-ion quantum computer. We successfully execute XY-QAOA circuits that restrict the quantum evolution to the in-constraint subspace, using up to 20 qubits and a two-qubit gate depth of up to 159. We demonstrate the necessity of directly encoding the constraints into the quantum circuit by showing the trade-off between the in-constraint probability and the quality of the solution that is implicit if unconstrained quantum optimization methods are used. We show that this trade-off makes choosing good parameters difficult in general. We compare XY-QAOA to the Layer Variational Quantum Eigensolver algorithm, which has a highly expressive constant-depth circuit, and the Quantum Approximate Optimization Algorithm. We discuss the respective trade-offs of the algorithms and implications for their execution on near-term quantum hardware.

97 MATHEMATICS AND COMPUTING↗

A Benthic Habitat Monitoring Approach for Marine and Hydrokinetic Sites (Final Technical Report)

This final technical report summarizes the work completed as part of the Standardized and Cost-Effective Benthic Habitat Mapping and Monitoring Tools for MHK Environmental Assessments project funded by U.S. Department of Energy (DOE) under contract DE-EE007826 to Integral Consulting Inc. The overall goal of this project was to demonstrate a consistent, repeatable, and semi-automated seafloor and sediment mapping approach for rapidly characterizing benthic physical and biological/habitat conditions to support marine environmental assessments for marine and hydrokinetic energy sites. The approach evaluated combines sediment profile imaging and plan view (SPI/PV) technology with multibeam echosounder surveys as an effective and low-cost benthic habitat mapping protocol. A key innovation was the development of a semi-automated computer vision system that standardizes the extraction of data from the SPI/PV images. Other elements of the project were SPI camera hardware modifications, including the design and fabrication of a prototype “power” SPI camera to improve camera prism penetration in firm substrates, and outreach to agency regulators and other stakeholders on this habitat mapping approach. This technical report consists of five main subsections that summarize: 1) the benthic mapping approach; 2) benthic mapping results from the three areas’ surveys; 3) the development and performance of the image processing algorithms; 4) SPI camera hardware improvements and prototype testing; and 5) the regulatory outreach efforts.

16 TIDAL AND WAVE POWER↗

K-Spin Hamiltonian for Quantum-Resolvable Markov Decision Processes

The Markov decision process is the mathematical formalization underlying the modern field of reinforcement learning when transition and reward functions are unknown. We derive a pseudo-Boolean cost function that is equivalent to a K-spin Hamiltonian representation of the discrete, finite, discounted Markov decision process with infinite horizon. This K-spin Hamiltonian furnishes a starting point from which to solve for an optimal policy using heuristic quantum algorithms such as adiabatic quantum annealing and the quantum approximate optimization algorithm on near-term quantum hardware. In arguing that the variational minimization of our Hamiltonian is approximately equivalent to the Bellman optimality condition for a prevalent class of environments we establish an interesting analogy with classical field theory. Along with proof-of-concept calculations to corroborate our formulation by simulated and quantum annealing against classical Q-Learning, we analyze the scaling of physical resources required to solve our Hamiltonian on quantum hardware.

Hamiltonian↗

Exploring the scaling limitations of the variational quantum eigensolver with the bond dissociation of hydride diatomic molecules

Abstract Materials simulations involving strongly correlated electrons pose fundamental challenges to state‐of‐the‐art electronic structure methods but are hypothesized to be the ideal use case for quantum computing algorithms. To date, no quantum computer has simulated a molecule of a size and complexity relevant to real‐world applications, despite the fact that the variational quantum eigensolver (VQE) algorithm can predict chemically accurate total energies. Nevertheless, because of the many applications of moderately sized, strongly correlated systems, such as molecular catalysts, the successful use of the VQE stands as an important waypoint in the advancement toward useful chemical modeling on near‐term quantum processors. In this paper, we take a significant step in this direction. We lay out the steps, write, and run parallel code for an (emulated) quantum computer to compute the bond dissociation curves of the TiH, LiH, NaH, and KH diatomic hydride molecules using the VQE. TiH was chosen as a relatively simple chemical system that incorporates d orbitals and strong electron correlation. Because current VQE implementations on existing quantum hardware are limited by qubit error rates, the number of qubits available, and the allowable gate depth, recent studies using it have focused on chemical systems involving s and p block elements. Through VQE + UCCSD calculations of TiH, we evaluate the near‐term feasibility of modeling a molecule with d‐orbitals on real quantum hardware. We demonstrate that the inclusion of d‐orbitals and the use of the UCCSD ansatz, which are both necessary to capture the correct TiH physics, dramatically increase the cost of this problem. We estimate the approximate error rates necessary to model TiH on current quantum computing hardware using VQE + UCCSD and show them to likely be prohibitive until significant improvements in hardware and error correction algorithms are available.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

The Design of a Fault-Tolerant COTS-Based Bus Architecture

In this paper, we report our experiences and findings on the design of a fault-tolerant bus architecture comprised of two COTS buses, the IEEE 1394 and the 12C. This fault-tolerant bus is the backbone system bus for the avionics architecture of the X2000 program at the Jet Propulsion Laboratory. COTS buses are attractive because of the availability of low cost commercial products. However, they are not specifically designed for highly reliable applications such as long-life deep-space missions. The X2000 design team has devised a multi-level fault tolerance approach to compensate for this shortcoming of COTS buses. First, the approach enhances the fault tolerance capabilities of the IEEE 1394 and 12 C buses by adding a layer of fault handling hardware and software. Second, algorithms are developed to enable the IEEE 1394 and the 12 C buses assist each other to isolate and recovery from faults. Third, the set of IEEE 1394 and 12 C buses is duplicated to further enhance system reliability. The X2000 design team has paid special attention to guarantee that all fault tolerance provisions will not cause the bus design to deviate from the commercial standard specifications. Otherwise, the economic attractiveness of using COTS will be diminished. The hardware and software design of the X2000 fault-tolerant bus are being implemented and flight hardware will be delivered to the ST4 and Europa Orbiter missions.

Chau, Savio N.↗

Analytical sensor redundancy assessment

The rationale and mechanization of sensor fault tolerance based on analytical redundancy principles are described. The concept involves the substitution of software procedures, such as an observer algorithm, to supplant additional hardware components. The observer synthesizes values of sensor states in lieu of their direct measurement. Such information can then be used, for example, to determine which of two disagreeing sensors is more correct, thus enhancing sensor fault survivability. Here a stability augmentation system is used as an example application, with required modifications being made to a quadruplex digital flight control system. The impact on software structure and the resultant revalidation effort are illustrated as well. Also, the use of an observer algorithm for wind gust filtering of the angle-of-attack sensor signal is presented.

Mulcare, D. B.↗

Quantum Computing to Accelerate High Fidelity Computational Materials Modeling

In this CIF we worked to develop quantum algorithms for material science simulations based on new ideas recently proposed on plane wave basis sets. Many body simulations are not generally performed in plane wave basis sets on classical hardware, but with newly proposed quantum algorithms, it is possible this will be a highly efficient basis set to run quantum simulations on quantum hardware. Using state of the art classical simulations we ran small test simulations to estimate the resources that will be needed to run such plane wave algorithms on quantum hardware. We demonstrate our approach for a series of atoms and molecular systems.

Norman Tubman↗

MRT 7365: Power flow physics and key physics phenomena

The Z accelerator at Sandia National Laboratories conducts z-pinch experiments at 26 MA in support of DOE missions in stockpile stewardship, dynamic materials, fusion, and other basic sciences. Increasing the current delivered to the z-pinch would extend our reach in each of these disciplines. To achieve increases in current and accelerator efficiency, a fraction of Z’s shots are set aside for research into transmission-line power flow. These shots, with supporting simulations and theory, are incorporated into this Advanced Diagnostics milestone report. The efficiency of Z is reduced as some portion of the total current is shunted across the transmission-line gaps prior to the load. This is referred to as “current loss”. Electrode plasmas have long been implicated in this process, so the bulk of dedicated power-flow experiments are designed to measure the plasma environment. The experimental analyses are enhanced by simulations conducted using realistic hardware and Z voltage pulses. In the same way that diagnostics are continually being improved for sensitivity and resolution, the modeling capability is continually being improved to provide faster and more realistic simulations. The specifics of the experimental hardware, diagnostics, simulations, and algorithm developments are provided in this report. The combined analysis of simulation and data confirms that electrode plasmas have the most detrimental impact on current delivery. Experiments over the last three years have tested the theoretical current-loss mechanisms of enhanced ion current, plasma gap closure, and Hall-related current. These mechanisms are not mutually exclusive and may be coincident in the final feed as well as in upstream transmission lines. The final-feed geometries tested here, however, observe lower-density plasmas without dominant ion currents which is consistent with a Hall-related current. The picture of plasma formation and transport formed from experiment and simulation is informing hardware designs being fielded on Z now and being proposed for the Next-Generation Pulsed Power (NGPP) facility. In this picture, the strong magnetic fields that heat the electrodes above particle emission thresholds also confine the charged particles near the surface. Some portion of the plasmas thus formed is transported into the transmission-line gap under the force of the electric field, with aid from plasma instabilities. The gap plasmas are then transported towards the load by a cross-field drift, where they accumulate and contribute to a likely Hall-related cross-gap current. The achievements in experimental execution, model validation, and physical analysis presented in this report set the stage for continued progress in power flow and load diagnostics on Z. The planned shot schedule for Z and Mykonos will provide data for extrapolation to higher current to ensure the predicted performance and efficiency of a NGPP facility.

43 PARTICLE ACCELERATORS↗

Horizon: A Proposal for Large Aperture, Active Optics in Geosynchronous Orbit

In 1999, NASA's New Millennium Program called for proposals to validate new technology in high-earth orbit for the Earth Observing-3 (NMP EO3) mission to fly in 2003. In response, we proposed to test a large aperture, active optics telescope in geosynchronous orbit. This would flight-qualify new technologies for both Earth and Space science: 1) a future instrument with LANDSAT image resolution and radiometric quality watching continuously from geosynchronous station, and 2) the Next Generation Space Telescope (NGST) for deep space imaging. Six enabling technologies were to be flight-qualified: 1) a 3-meter, lightweight segmented primary mirror, 2) mirror actuators and mechanisms, 3) a deformable mirror, 4) coarse phasing techniques, 5) phase retrieval for wavefront control during stellar viewing, and 6) phase diversity for wavefront control during Earth viewing. Three enhancing technologies were to be flight- validated: 1) mirror deployment and latching mechanisms, 2) an advanced microcontroller, and 3) GPS at GEO. In particular, two wavefront sensing algorithms, phase retrieval by JPL and phase diversity by ERIM International, were to sense optical system alignment and focus errors, and to correct them using high-precision mirror mechanisms. Active corrections based on Earth scenes are challenging because phase diversity images must be collected from extended, dynamically changing scenes. In addition, an Earth-facing telescope in GEO orbit is subject to a powerful diurnal thermal and radiometric cycle not experienced by deep-space astronomy. The Horizon proposal was a bare-bones design for a lightweight large-aperture, active optical system that is a practical blend of science requirements, emerging technologies, budget constraints, launch vehicle considerations, orbital mechanics, optical hardware, phase-determination algorithms, communication strategy, computational burdens, and first-rate cooperation among earth and space scientists, engineers and managers. This manuscript presents excerpts from the Horizon proposal's sections that describe the Earth science requirements, the structural -thermal-optical design, the wavefront sensing and control, and the on-orbit validation.

Chesters, Dennis↗

tih_vqe [SWR-23-32]

This software supports the paper, "Exploring the scaling limitations of the variational quantum eigensolver with the bond dissociation of hydride diatomic molecules," published in the International Journal of Quantum Chemistry, whose abstract is as follows: Materials simulations involving strongly correlated electrons pose fundamental challenges to state-of-the-art electronic structure methods but are hypothesized to be the ideal use case for quantum computing. To date, no quantum computer has simulated a molecule of a size and complexity relevant to real-world applications, despite the fact that the variational quantum eigensolver (VQE) algorithm can predict chemically accurate total energies. Nevertheless, because of the many applications of moderately-sized, strongly correlated systems, such as molecular catalysts, the successful use of the VQE stands as an important waypoint in the advancement toward useful chemical modeling on near-term quantum processors. In this paper, we take a significant step in this direction. We lay out the steps, write, and run parallel code for an (emulated) quantum computer to compute the bond dissociation curves of the TiH, LiH, NaH, and KH diatomic hydride molecules using VQE. TiH was chosen as a relatively simple chemical system that incorporates d orbitals and strong electron correlation. Because current VQE implementations on existing quantum hardware are limited by qubit error rates, the number of qubits available, and the allowable gate depth, recent studies have focused on chemical systems involving s and p block elements. Through VQE + UCCSD calculations of TiH, we evaluate the near-term feasibility of modeling a molecule with d-orbitals on real quantum hardware. We demonstrate that the inclusion of d-orbitals and the use of the UCCSD ansatz, which are both necessary to capture the correct TiH physics, dramatically increase the cost of this problem. We estimate the approximate error rates necessary to model TiH on current quantum computing hardware using VQE+UCCSD and show them to likely be prohibitive until significant improvements in hardware and error correction algorithms are available.

Graf, Peter↗

A Comparison of PETSC Library and HPF Implementations of an Archetypal PDE Computation

Two paradigms for distributed-memory parallel computation that free the application programmer from the details of message passing are compared for an archetypal structured scientific computation a nonlinear, structured-grid partial differential equation boundary value problem using the same algorithm on the same hardware. Both paradigms, parallel libraries represented by Argonne's PETSC, and parallel languages represented by the Portland Group's HPF, are found to be easy to use for this problem class, and both are reasonably effective in exploiting concurrency after a short learning curve. The level of involvement required by the application programmer under either paradigm includes specification of the data partitioning (corresponding to a geometrically simple decomposition of the domain of the PDE). Programming in SPAM style for the PETSC library requires writing the routines that discretize the PDE and its Jacobian, managing subdomain-to-processor mappings (affine global- to-local index mappings), and interfacing to library solver routines. Programming for HPF requires a complete sequential implementation of the same algorithm, introducing concurrency through subdomain blocking (an effort similar to the index mapping), and modest experimentation with rewriting loops to elucidate to the compiler the latent concurrency. Correctness and scalability are cross-validated on up to 32 nodes of an IBM SP2.

Hayder, M. Ehtesham↗

Performance Evaluation of Intelligent Solar Control Software Through Hardware-in-the-Loop (CRADA Final Report)

Recent research has highlighted the potential for solar to act as a zero-marginal-cost and zero-emission flexibility resource on the bulk power system when operated with advanced control systems. To increase the performance of these systems, leading technologies, including machine learning (ML) and hierarchical inverter set point allocation, have been developed by Latimer Controls, Inc. to estimate the headroom of large PV plants for grid operation and control; however, these technologies lack comprehensive validation under real-world application scenarios. Latimer Controls, Inc. received two voucher awards for research at a national laboratory from the Department of Energy American Made Solar Prize Round 6. The National Renewable Energy Laboratory (NREL) was selected to collaborate with Latimer staff to conduct a performance evaluation of Latimer PV control software. The NREL team will develop a hardware-in-the-loop (HIL) testbed to perform testing and validation of the Latimer PV control technology in a de-risked yet realistic testbed environment. Latimer and NREL worked together to analyze the test data, draw conclusions from the results, and disseminate the resulting scientific findings. In this CRADA work, we propose to test and validate the real-world application of the Latimer Control solution in an HIL environment. We evaluate the performance of different flexible solar technologies in responding to automatic generation control signals in a closed-loop fashion. In particular, a data-driven potential high limit (PHL) estimation is developed for large solar plants to accurately estimate their headroom so that they have fast and short-time regulation and control capability to participate in grid services and respond to grid signals in real time (e.g., AGC). This PHL estimation algorithm is embedded in a hardware power plant controller (PPC) and tested with an IEEE-39 bus system model developed in RTDS. To account for the varying cloud conditions and diverse inverter dispatches, we developed a 135-MW PV plant with detailed modeling of 27 individual PV modules and inverters using RTDS. The real-world communications used in such big plants, such as ModBus TCP/IP for inverter level and DNP3 for plant level, were developed to emulate the real-world applications in big PV plants. The ML-based PHL estimation method is tested under nine separate weather scenarios against the ‘reference-control’ solution, hereafter referred to as the baseline solution. The baseline method reserves a subset of inverters (reference group) to operate at their PHL at all times and dispatches only the remaining inverters (control group) at curtailed levels to fulfill the flexibility need. Despite being successfully piloted by NREL in California in 2017 and Chile in 2020, there exist two gaps in the state of the art to fully unlock the flexibility of PV plants: a. There is a trade-off between the PHL estimation accuracy and the flexibility range. b. There lacks granularity in the PHL estimation to capture the variation across inverters. The Latimer solution seeks to address these gaps by applying machine learning methods to improve PHL estimation accuracy while accounting for variability at every inverter. Performance metrics were taken from the 2023 Georgia Power CARES utility-scale RFP. The results demonstrate that the ML-based approach outperforms the traditional baseline method in PHL estimation accuracy for 7 of 9 scenarios. The average PHL error across the nine scenarios was 7.40% for the ML-based method, 2.06% less than the 9.46% PHL error average across scenarios that was exhibited by the baseline method. Additionally, the PHL error was below 5% for at least 95% of the testing interval for 3 of 9 tested intervals with the ML approach, whereas it did not achieve this metric for any of the baseline tests. Overall, simulation results indicate the superior performance of an ML-based approach compared to the conventional baseline reference-control approach, showcasing its potential to support grid stability and operational efficiency. This laboratory HIL testing using real PPC, representative power system simulation models in real-time with detailed PV plant and inverter models, and real-world communication protocols gives us confidence that this machine learning based PHL estimation algorithm works well in the hardware PPC and therefore de-risks future field commissioning. The end goal of this project is to advance grid technology to address the grid operation challenges brought by solar plant’s variability and uncertainties in power generation.

14 SOLAR ENERGY↗

Flexible silicon photonic architecture for accelerating distributed deep learning

The increasing size and complexity of deep learning (DL) models have led to the wide adoption of distributed training methods in datacenters (DCs) and high-performance computing (HPC) systems. However, communication among distributed computing units (CUs) has emerged as a major bottleneck in the training process. In this study, we propose Flex-SiPAC, a flexible silicon photonic accelerated compute cluster designed to accelerate multi-tenant distributed DL training workloads. Flex-SiPAC takes a co-design approach that combines a silicon photonic hardware platform with a tailored collective algorithm, optimized to leverage the unique physical properties of the architecture. The hardware platform integrates a novel wavelength-reconfigurable transceiver design and a micro-resonator-based wavelength-reconfigurable switch, enabling the system to achieve flexible bandwidth steering in the wavelength domain. The collective algorithm is designed to support reconfigurable topologies, enabling efficient all-reduce communications that are commonly used in DL training. The feasibility of the Flex-SiPAC architecture is demonstrated through two testbed experiments. First, an optical testbed experiment demonstrates the flexible routing of wavelengths by shuffling an array of input wavelengths using a custom-designed spatial-wavelength selective switch. Second, a four-GPU testbed running two DL workloads shows a 23% improvement in job completion time compared to a similarly sized leaf-spine topology. We further evaluate Flex-SiPAC using large-scale simulations, which show that Flex-SiPAC is able to reduce the communication time by 26% to 29% compared to state-of-the-art compute clusters under representative collective operations.

Wu, Zhenguo (ORCID:0000000322847985)↗

Hybrid-Electric Aero-Propulsion Controls Testbed Results with Energy Storage

Electrified aircraft propulsion (EAP) research is a priority of the National Aeronautics and Space Administration (NASA) for its potential to increase propulsion system efficiency, performance, and operability at the subsystem and vehicle levels while decreasing emissions. These EAP systems demand more advanced control algorithms due to increased complexity. NASA has developed a reconfigurable, hardware-in-the-loop rig to verify control algorithm performance using a sub-scale electro-mechanical system. A novel capability of this rig is the ability to test full scale EAP control algorithms on a sub-scale representation of the electro-mechanical system without turbomachinery/rotors. A novel feature is the use of a physical energy storage device within the electro-mechanical system. A dual spool, parallel hybrid-electric turbofan architecture and energy management control system is tested with the goal of verifying the ability to obtain turbomachinery model operability benefits while controlling sub-scale electro-mechanical hardware. Pre-test predictions of the turbofan model, control, and rig performance were obtained through simulation using a software model of the rig. Theoretical results showing the true performance of the turbofan model were obtained through a software simulation using full-scale mechanical shaft models. The paper compares theoretical, predicted, and actual test results from the turbofan model, energy management control and rig perspectives. The results show that the presence of sub-scale electro-mechanical hardware did not inhibit the energy management algorithm from achieving turbomachinery operability benefits.

hybrid↗

Hybrid-Electric Aero-Propulsion Controls Testbed Results with Energy Storage

Electrified aircraft propulsion (EAP) research is a priority of the National Aeronautics and Space Administration (NASA) for its potential to increase propulsion system efficiency, performance, and operability at the subsystem and vehicle levels while decreasing emissions. These EAP systems demand more advanced control algorithms due to increased complexity. NASA has developed a reconfigurable, hardware-in-the-loop rig to verify control algorithm performance using a sub-scale electro-mechanical system. A novel capability of this rig is the ability to test full scale EAP control algorithms on a sub-scale representation of the electro-mechanical system without turbomachinery/rotors. A novel feature is the use of a physical energy storage device within the electro-mechanical system. A dual spool, parallel hybrid-electric turbofan architecture and energy management control system is tested with the goal of verifying the ability to obtain turbomachinery model operability benefits while controlling sub-scale electro-mechanical hardware. Pre-test predictions of the turbofan model, control, and rig performance were obtained through simulation using a software model of the rig. Theoretical results showing the true performance of the turbofan model were obtained through a software simulation using full-scale mechanical shaft models. The paper compares theoretical, predicted, and actual test results from the turbofan model, energy management control and rig perspectives. The results show that the presence of sub-scale electro-mechanical hardware did not inhibit the energy management algorithm from achieving turbomachinery operability benefits.

hybrid↗

VHP - An environment for the remote visualization of heuristic processes

A software system called VHP is introduced which permits the visualization of heuristic algorithms on both resident and remote hardware platforms. The VHP is based on the DCF tool for interprocess communication and is applicable to remote algorithms which can be on different types of hardware and in languages other than VHP. The VHP system is of particular interest to systems in which the visualization of remote processes is required such as robotics for telescience applications.

Crawford, Stuart L.↗

Implementation of real-time digital signal processing systems

Special purpose hardware implementation of DFT Computers and digital filters is considered in the light of newly introduced algorithms and IC devices. Recent work by Winograd on high-speed convolution techniques for computing short length DFT's, has motivated the development of more efficient algorithms, compared to the FFT, for evaluating the transform of longer sequences. Among these, prime factor algorithms appear suitable for special purpose hardware implementations. Architectural considerations in designing DFT computers based on these algorithms are discussed. With the availability of monolithic multiplier-accumulators, a direct implementation of IIR and FIR filters, using random access memories in place of shift registers, appears attractive. The memory addressing scheme involved in such implementations is discussed. A simple counter set-up to address the data memory in the realization of FIR filters is also described. The combination of a set of simple filters (weighting network) and a DFT computer is shown to realize a bank of uniform bandpass filters. The usefulness of this concept in arriving at a modular design for a million channel spectrum analyzer, based on microprocessors, is discussed.

Narasimha, M.↗