Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Hardware generation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 235 records · Page 13

Airborne Optical Communications Demonstrator Design And Preflight Test Results

A second generation optical communications demonstrator (OCD-2) intended for airborne applications like air-to-ground and air-to-air optical links is under development at JPL. This development provides the capability for unidirectional high data rate (2.5-Gbps) transmission at 1550-nm, with the ability to receive an 810-nm beacon to aid acquisition pointing and tracking. The transmitted beam width is nominally 200-(micro)rad. A 3x3 degree coarse field-of-view (FOV) acquisition sensor with a much smaller ~3-mrad FOV tracking sensor is incorporated. The OCD-2 optical head will be integrated to a high performance gimbal turret assembly capable of providing pointing stability of 5- microradians from an airborne platform. Other parts of OCD-2 include a cable harness, connecting the optical head in the gimbal turret assembly to a rugged electronics box. The electronics box will house: command and control processors, laser transmitter, data-generation-electronics, power conversion/distribution hardware and state-of-health monitors. The entire assembly will be integrated and laboratory tested prior to a planned flight demonstrations.

optical communications↗

Analysis of Large Civil Tilt Rotor Wind Tunnel Blockage and Validation Using RotCFD

Ground based experiments are often used to understand and measure rotor and airframe aerodynamic performance; however, these experiments have certain limitations. The effects of these limitations are evaluated here using computational fluid dynamic (CFD) modeling techniques. Through this study, data from the 7- by 10-Foot Wind Tunnel experiments of the Large Civil Tilt Rotor (LCTR) at NASA Ames Research Center is validated using CFD. The Reynolds Averages Navier-Stokes solver, RotCFD, is used for the computations. In particular, the effect of the blockage generated by the test hardware on the walls is investigated. To study this problem, simplified geometries such as a flat plate, cube and cylinder are also investigated for blockage effects. This is done to explore if these different geometries can represent the LCTR as a simplified case to reduce computational time and get a quick first understanding of tunnel blockage effects. The focus of this research is to understand the limitations and accuracy of the recent small-scale Large Civil Tilt Rotor wind tunnel test campaigns.

Validation↗

The Power of Pausing: Advancing Understanding of Thermalization in Experimental Quantum Annealers

We investigate alternative annealing schedules on the current generation of quantum annealing hardware (the D-Wave 2000Q), which includes the use of forward and reverse annealing with an- intermediate pause. This work provides new insights into the inner workings of these devices (and quantum devices in general), particular into how thermal effects govern the system dynamics. We show that a pause mid-way through the anneal can cause a dramatic change in the output distribution, and we provide evidence suggesting thermalization is indeed occurring during such a pause. We demonstrate that upon pausing the system in a narrow region shortly after the minimum gap, the probability of successfully finding the ground state of the problem Hamiltonian can be increased by several orders of magnitude. We relate this effect to relaxation (i.e. thermalization) after diabatic and thermal excitations that occur in the region near to the minimum gap. For a set of large-scale problems of up to 500 qubits, we demonstrate that the distribution returned from the annealer very closely matches a (classical) Boltzmann distribution of the problem Hamiltonian, albeit one with a temperature at least 1.5 times higher than the (effective) temperature of the device. Moreover, we show that larger problems are more likely to thermalize to a classical Boltzmann distribution.

Marshall, Jeffrey↗

NASA's Development of Advanced Space Suits for Lunar Exploration

Significant work has been completed recently at NASA’s Johnson Space Center in the Space Suit and Crew Survival Systems Branch to design and mature an advanced exploration EVA (Extra-Vehicular Activity) architecture. This effort culminated in the xEMU (eXploration Extra-vehicular Mobility Unit) space suit, which was completed in 2022 and tested extensively through 2023, including human performance, cycle life performance and thermal vacuum testing. An overview of the xEMU design and test campaign is provided, as well as its value as the government reference design supporting the transition to EVA commercial services for the Artemis program. Lastly, a development roadmap is discussed which outlines NASA’s identified strategy to enable the next generation of exploration EVA hardware for sustaining-class Lunar and Mars missions.

Shane M McFarland↗

Per-instruction energy debugging using instruction sampling hardware

A processor utilizes instruction based sampling to generate sampling data sampled on a per instruction basis during execution of an instruction. The sampling data indicates what processor hardware was used due to the execution of the instruction. Software receives the sampling data and generates an estimate of energy used by the instruction based on the sampling data. The sampling data may include microarchitectural events and the energy estimate utilizes a base energy amount corresponding to the instruction executed along with energy amounts corresponding to the microarchitectural events in the sampling data. The sampling data may include switching events associated with hardware blocks that switched due to execution of the instruction and the energy estimate for the instruction is based on the switching events and capacitance estimates associated with the hardware blocks.

Wei, Shijia↗

miniGAN: A Generative Adversarial Network proxy application WBS 2.2.6.08 ECP-2.1.3 (Q3 FY2020 Milestone Report) (V.1.0)

In order to support the machine learning co-design needs of ECP applications in current and future DOE HPC hardware, we have developed a generative adversarial network (GAN) proxy application, miniGAN, that has been released through the ECP proxy application suite. The proxy application is representative of the needs of ExaLearn's target applications, specifically the Cosmoflow and ExaGAN cosmology applications and the ExaWind energy application. The proxy application also demonstrates the first use of performance portable kernels within widely-used machine learning frameworks: PyTorch (Facebook) and Horovod (Uber). We provide performance scaling results for similar workloads to ExaGAN and a profile of individual GAN training components.

97 MATHEMATICS AND COMPUTING↗

Independent Orbiter Assessment (IOA): Analysis of the electrical power generation/fuel cell powerplant subsystem

Results of the Independent Orbiter Assessment (IOA) of the Failure Modes and Effects Analysis (FMEA) and Critical Items List (CIL) are presented. The IOA approach features a top-down analysis of the hardware to determine failure modes, criticality, and potential critical items. To preserve independence, this analysis was accomplished without reliance upon the results contained within the NASA FMEA/CIL documentation. This report documents the independent analysis results corresponding to the Orbiter Electrical Power Generation (EPG)/Fuel Cell Powerplant (FCP) hardware. The EPG/FCP hardware is required for performing functions of electrical power generation and product water distribution in the Orbiter. Specifically, the EPG/FCP hardware consists of the following divisions: (1) Power Section Assembly (PSA); (2) Reactant Control Subsystem (RCS); (3) Thermal Control Subsystem (TCS); and (4) Water Removal Subsystem (WRS). The IOA analysis process utilized available EPG/FCP hardware drawings and schematics for defining hardware assemblies, components, and hardware items. Each level of hardware was evaluated and analyzed for possible failure modes and effects. Criticality was assigned based upon the severity of the effect for each failure mode.

Brown, K. L.↗

Intrinsic Hardware Evolution for the Design and Reconfiguration of Analog Speed Controllers for a DC Motor

Evolvable hardware provides the capability to evolve analog circuits to produce amplifier and filter functions. Conventional analog controller designs employ these same functions. Analog controllers for the control of the shaft speed of a DC motor are evolved on an evolvable hardware platform utilizing a second generation Field Programmable Transistor Array (FPTA2). The performance of an evolved controller is compared to that of a conventional proportional-integral (PI) controller. It is shown that hardware evolution is able to create a compact design that provides good performance, while using considerably less functional electronic components than the conventional design. Additionally, the use of hardware evolution to provide fault tolerance by reconfiguring the design is explored. Experimental results are presented showing that significant recovery of capability can be made in the face of damaging induced faults.

Gwaltney, David A.↗

Issues in designing transport layer multicast facilities

Multicasting denotes a facility in a communications system for providing efficient delivery from a message's source to some well-defined set of locations using a single logical address. While modem network hardware supports multidestination delivery, first generation Transport Layer protocols (e.g., the DoD Transmission Control Protocol (TCP) (15) and ISO TP-4 (41)) did not anticipate the changes over the past decade in underlying network hardware, transmission speeds, and communication patterns that have enabled and driven the interest in reliable multicast. Much recent research has focused on integrating the underlying hardware multicast capability with the reliable services of Transport Layer protocols. Here, we explore the communication issues surrounding the design of such a reliable multicast mechanism. Approaches and solutions from the literature are discussed, and four experimental Transport Layer protocols that incorporate reliable multicast are examined.

Dempsey, Bert J.↗

FFTX-IRIS: Towards Performance Portability and Heterogeneity for SPIRAL Generated Code

FFTX-IRIS is a dynamic system to efficiently utilize novel heterogeneous platforms. This system links two next-generation frameworks, FFTX and IRIS, to navigate the complexity of different hardware architectures. FFTX provides a runtime code generation framework for high-performance Fast Fourier Transform kernels. IRIS runtime provides portability and multi-device heterogeneity, allowing computation on any available compute resource. Together, FFTX-IRIS enables code generation, seamless portability, and performance without user involvement. We show the design of the FFTX-IRIS system along with an evaluation of various small FFT benchmarks. We also demonstrate multi-device heterogeneity of FFTX-IRIS with a larger stencil application.

Rao, Sanil↗

Hyperdimensional computing for image classification (HDC) v1.0

This is an implementation of the hyperdimensional computing technique to classify images. It consists of a python script that trains the system for a set of images from a set of images (dataset) specified by the user. This training produces hardware configuration parameters and description vectors that are then loaded into the hardware description part of the project. The hardware description consists of hardware described in Verilog (a well known language for this purpose) that is synthesizable and can be implemented in a real chip. This hardware received the training information generated by python, and then is able to accept images to produce answers for each image on which category (class) from the pre-=trained ones the image belongs to. The hardware and python training scripts are configurable and documented. The advantage of hyperdimensional computing is its robustness to errors and the easy capability for online learning (refining the training during inference slowly over time), which this implementation supports.

Michelogiannakis, Georgios [Lawrence Berkeley Nati↗

Final Report for Intravenous Fluid Generation (IVGEN) Spaceflight Experiment

NASA designed and operated the Intravenous Fluid Generation (IVGEN) experiment onboard the International Space Station (ISS), Increment 23/24, during May 2010. This hardware was a demonstration experiment to generate intravenous (IV) fluid from ISS Water Processing Assembly (WPA) potable water using a water purification technique and pharmaceutical mixing system. The IVGEN experiment utilizes a deionizing resin bed to remove contaminants from feedstock water to a purity level that meets the standards of the United States Pharmacopeia (USP), the governing body for pharmaceuticals in the United States. The water was then introduced into an IV bag where the fluid was mixed with USP-grade crystalline salt to produce USP normal saline (NS). Inline conductivity sensors quantified the feedstock water quality, output water purity, and NS mixing uniformity. Six 1.5-L bags of purified water were produced. Two of these bags were mixed with sodium chloride to make 0.9 percent NS solution. These two bags were returned to Earth to test for compliance with USP requirements. On-orbit results indicated that all of the experimental success criteria were met with the exception of the salt concentration. Problems with a large air bubble in the first bag of purified water resulted in a slightly concentrated saline solution of 117 percent of the target value of 0.9 g/L. The second bag had an inadequate amount of salt premeasured into the mixing bag resulting in a slightly deficient salt concentration of 93.8 percent of the target value. The USP permits a range from 95 to 105 percent of the target value. The testing plans for improvements for an operational system are also presented.

McQuillen, John B.↗

RotCFD Analysis of the AH-56 Cheyenne Hub Drag

In 2016, the U.S. Army Aviation Development Directorate (ADD) conducted tests in the U.S. Army 7- by 10- Foot Wind Tunnel at NASA Ames Research Center of a nonrotating 2/5th-scale AH-56 rotor hub. The objective of the tests was to determine how removing the mechanical control gyro affected the drag. Data for the lift, drag, and pitching moment were recorded for the 4-bladed rotor hub in various hardware configurations, azimuth angles, and angles of attack. Numerical simulations of a selection of the configurations and orientations were then performed, and the results were compared with the test data. To generate the simulation results, the hardware configurations were modeled using Creo and Rhinoceros 5, three-dimensional surface modeling computer-aided design (CAD) programs. The CAD model was imported into Rotorcraft Computational Fluid Dynamics (RotCFD), a computational fluid dynamics (CFD) tool used for analyzing rotor flow fields. RotCFD simulation results were compared with the experimental results of three hardware configurations at two azimuth angles, two angles of attack, and with and without wind tunnel walls. The results help validate RotCFD as a tool for analyzing low-drag rotor hub designs for advanced high-speed rotorcraft concepts. Future work will involve simulating additional hub geometries to reduce drag or tailor to other desired performance levels.

AH-56 Cheyenne↗

RenderMan design principles

The two worlds of interactive graphics and realistic graphics have remained separate. Fast graphics hardware runs simple algorithms and generates simple looking images. Photorealistic image synthesis software runs slowly on large expensive computers. The time has come for these two branches of computer graphics to merge. The speed and expense of graphics hardware is no longer the barrier to the wide acceptance of photorealism. There is every reason to believe that high quality image synthesis will become a standard capability of every graphics machine, from superworkstation to personal computer. The significant barrier has been the lack of a common language, an agreed-upon set of terms and conditions, for 3-D modeling systems to talk to 3-D rendering systems for computing an accurate rendition of that scene. Pixar has introduced RenderMan to serve as that common language. RenderMan, specifically the extensibility it offers in shading calculations, is discussed.

Apodaca, Tony↗

Performance Measurement of Advanced Stirling Convertors (ASC-E3)

NASA Glenn Research Center (GRC) has been supporting development of the Advanced Stirling Radioisotope Generator (ASRG) since 2006. A key element of the ASRG project is providing life, reliability, and performance testing data of the Advanced Stirling Convertor (ASC). The latest version of the ASC (ASC-E3, to represent the third cycle of engineering model test hardware) is of a design identical to the forthcoming flight convertors. For this generation of hardware, a joint Sunpower and GRC effort was initiated to improve and standardize the test support hardware. After this effort was completed, the first pair of ASC-E3 units was produced by Sunpower and then delivered to GRC in December 2012. GRC has begun operation of these units. This process included performance verification, which examined the data from various tests to validate the convertor performance to the product specification. Other tests included detailed performance mapping that encompassed the wide range of operating conditions that will exist during a mission. These convertors were then transferred to Lockheed Martin for controller checkout testing. The results of this latest convertor performance verification activity are summarized here.

Stirling Cycle↗

DSS-13 26-meter antenna upgraded radiometer system

The Deep Space Station (DSS)-13 26-m antenna radiometer system was upgraded with an IBM-compatible computer-controlled configuration with improved supporting hardware and software. Software was generated to analyze results and correct for antenna mispointing, tropospheric loss, and other observing errors. This total power radiometer configuration provides a prototype for the new DSS-13 34-m antenna. The radiometer system is described in terms of the theory, instrumentation hardware, computer configuration, and operational features and performance. The system is used to obtain antenna efficiency and pointing model data and is useful for radio source calibrations required for radio astronomy. Some recent results are given.

Stelzried, C. T.↗

autoGEMM: Pushing the Limits of Irregular Matrix Multiplication on Arm Architectures

This paper presents an open-source library that pushes the limits of performance portability for irregular General Matrix Multiplication (GEMM) on the widely-used Arm architectures. Our library, autoGEMM, is designed to support a wide range of Arm processors: from edge devices to HPC-grade CPUs. autoGEMM generates optimized kernels for various hardware configurations by auto-combining fragments of autogenerated micro-kernels that employ hand-written optimizations to maximize computational efficiency. We optimize the kernel pipeline by tuning the register reuse and the data load/store overlapping. In addition, we use a dynamic tiling scheme to generate balanced tile shapes. Finally, we position autoGEMM on top of the TVM framework where our dynamic tiling scheme prunes the search space for TVM to identify the optimal combination of parameters for code optimization. Evaluations on five different classes of Arm chips demonstrate the advantages of autoGEMM. For small matrices, autoGEMM achieves 98% of peak and up to 2.0x speedup over state-of-the-art libraries such as LIBXSMM and LibShalom. For irregular matrices (i.e. tall skinny and long rectangles), autoGEMM is 1.3-2.0x faster than widely-used libraries such as OpenBLAS and Eigen. autoGEMM is publicly available at: https://github.com/wudu98/autoGEMM.

Wu, Du↗

Real-Time Krylov Theory for Quantum Computing Algorithms

Quantum computers provide new avenues to access ground and excited state properties of systems otherwise difficult to simulate on classical hardware. New approaches using subspaces generated by real-time evolution have shown efficiency in extracting eigenstate information, but the full capabilities of such approaches are still not understood. In recent work, we developed the variational quantum phase estimation (VQPE) method, a compact and efficient real-time algorithm to extract eigenvalues on quantum hardware. Here we build on that work by theoretically and numerically exploring a generalized Krylov scheme where the Krylov subspace is constructed through a parametrized real-time evolution, which applies to the VQPE algorithm as well as others. We establish an error bound that justifies the fast convergence of our spectral approximation. We also derive how the overlap with high energy eigenstates becomes suppressed from real-time subspace diagonalization and we visualize the process that shows the signature phase cancellations at specific eigenenergies. We investigate various algorithm implementations and consider performance when stochasticity is added to the target Hamiltonian in the form of spectral statistics. To demonstrate the practicality of such real-time evolution, we discuss its application to fundamental problems in quantum computation such as electronic structure predictions for strongly correlated systems.

97 MATHEMATICS AND COMPUTING↗