Engineering PapersSearch

SEARCH · Engineering Papers

Results for “HPC”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Computational Aspects of Data Assimilation and the ESMF

The scientific challenge of developing advanced data assimilation applications is a daunting task. Independently developed components may have incompatible interfaces or may be written in different computer languages. The high-performance computer (HPC) platforms required by numerically intensive Earth system applications are complex, varied, rapidly evolving and multi-part systems themselves. Since the market for high-end platforms is relatively small, there is little robust middleware available to buffer the modeler from the difficulties of HPC programming. To complicate matters further, the collaborations required to develop large Earth system applications often span initiatives, institutions and agencies, involve geoscience, software engineering, and computer science communities, and cross national borders.The Earth System Modeling Framework (ESMF) project is a concerted response to these challenges. Its goal is to increase software reuse, interoperability, ease of use and performance in Earth system models through the use of a common software framework, developed in an open manner by leaders in the modeling community. The ESMF addresses the technical and to some extent the cultural - aspects of Earth system modeling, laying the groundwork for addressing the more difficult scientific aspects, such as the physical compatibility of components, in the future. In this talk we will discuss the general philosophy and architecture of the ESMF, focussing on those capabilities useful for developing advanced data assimilation applications.

daSilva, A.

Effect of Exercise Training and +Gz Acceleration Training on Men

Countermeasures for reduction in work capacity (maximal oxygen uptake and strength) during spaceflight and enhanced orthostatic intolerance during re-entry, landing and egress from the return vehicle are continuing problems. The purpose for this study was to test the hypothesis that passive-acceleration training; supine, interval, exercise plus acceleration training and exercise combined with acceleration training would improve orthostatic tolerance in ambulatory men; and that addition of the aerobic exercise conditioning would not alter this improved tolerance from that of passive-acceleration training. Seven men (24-38 yr) underwent "Passive" training on the Ames human-powered centrifuge (HPC) for 30 min, "Exercise" training on the cycle ergometer with constant +Gz acceleration; and "Combined" exercise training at 40% to 90% of the HPC +Gz(max) exercise level. Maximal supine exercise loads increased significant (P<0.05) by 8.3% (Passive), 12.6% (Exercise), and by 15.4% (Combined) after training, but their post-training maximal oxygen uptakes and maximal heart rates were unchanged. Maximal time to fatigue (endurance) was unchanged with Passive was increased (P<0.05) with Exercise and Combined training. Thus, the exercise in the Exercise and Combined training Phases resulted in greater maximal loads and endurance without effect on maximal oxygen uptake or heart rate. There was a 4% to 6% increase (P<0.05) in all four quadriceps muscle volumes (right and left) after post-Combined training. Resting pre-tilt heart rate was elevated by 12.9% (P<0.05) only after Passive training suggesting that the exercise training attenuated the HR response. Plasma volume (% Delta) was uniformly decreased by 8% to 14% (P<0.05) at tilt-tolerance pre- vs. post-training indicating essentially no effect of training on the level of hypovolemia. Post-training tilt-tolerance time and heart rate were increased (P<0.05) only with Passive training by 37.8% and by 29.1%, respectively. Thus, addition of exercise training appeared to attenuate the increased Passive tilt-tolerance.

Greenleaf, John E.

Multi-Core Processor Memory Contention Benchmark Analysis Case Study

Multi-core processors dominate current mainframe, server, and high performance computing (HPC) systems. This paper provides synthetic kernel and natural benchmark results from an HPC system at the NASA Goddard Space Flight Center that illustrate the performance impacts of multi-core (dual- and quad-core) vs. single core processor systems. Analysis of processor design, application source code, and synthetic and natural test results all indicate that multi-core processors can suffer from significant memory subsystem contention compared to similar single-core processors.

Simon, Tyler

Constructive Engineering of Simulations

Joint experimentation that investigates sensor optimization, re-tasking and management has far reaching implications for Department of Defense, Interagency and multinational partners. An adaption of traditional human in the loop (HITL) Modeling and Simulation (M&S) was one approach used to generate the findings necessary to derive and support these implications. Here an entity-based simulation was re-engineered to run on USJFCOM's High Performance Computer (HPC). The HPC was used to support the vast number of constructive runs necessary to produce statistically significant data in a timely manner. Then from the resulting sensitivity analysis, event designers blended the necessary visualization and decision making components into a synthetic environment for the HITL simulations trials. These trials focused on areas where human decision making had the greatest impact on the sensor investigations. Thus, this paper discusses how re-engineering existing M&S for constructive applications can positively influence the design of an associated HITL experiment.

Snyder, Daniel R.

Scheduling Operations for Massive Heterogeneous Clusters

High-performance computing (HPC) programming has become increasingly difficult with the advent of hybrid supercomputers consisting of multicore CPUs and accelerator boards such as the GPU. Manual tuning of software to achieve high performance on this type of machine has been performed by programmers. This is needlessly difficult and prone to being invalidated by new hardware, new software, or changes in the underlying code. A system was developed for task-based representation of programs, which when coupled with a scheduler and runtime system, allows for many benefits, including higher performance and utilization of computational resources, easier programming and porting, and adaptations of code during runtime. The system consists of a method of representing computer algorithms as a series of data-dependent tasks. The series forms a graph, which can be scheduled for execution on many nodes of a supercomputer efficiently by a computer algorithm. The schedule is executed by a dispatch component, which is tailored to understand all of the hardware types that may be available within the system. The scheduler is informed by a cluster mapping tool, which generates a topology of available resources and their strengths and communication costs. Software is decoupled from its hardware, which aids in porting to future architectures. A computer algorithm schedules all operations, which for systems of high complexity (i.e., most NASA codes), cannot be performed optimally by a human. The system aids in reducing repetitive code, such as communication code, and aids in the reduction of redundant code across projects. It adds new features to code automatically, such as recovering from a lost node or the ability to modify the code while running. In this project, the innovators at the time of this reporting intend to develop two distinct technologies that build upon each other and both of which serve as building blocks for more efficient HPC usage. First is the scheduling and dynamic execution framework, and the second is scalable linear algebra libraries that are built directly on the former.

Humphrey, John

Near Real-Time Probabilistic Damage Diagnosis Using Surrogate Modeling and High Performance Computing

This work investigates novel approaches to probabilistic damage diagnosis that utilize surrogate modeling and high performance computing (HPC) to achieve substantial computational speedup. Motivated by Digital Twin, a structural health management (SHM) paradigm that integrates vehicle-specific characteristics with continual in-situ damage diagnosis and prognosis, the methods studied herein yield near real-time damage assessments that could enable monitoring of a vehicle's health while it is operating (i.e. online SHM). High-fidelity modeling and uncertainty quantification (UQ), both critical to Digital Twin, are incorporated using finite element method simulations and Bayesian inference, respectively. The crux of the proposed Bayesian diagnosis methods, however, is the reformulation of the numerical sampling algorithms (e.g. Markov chain Monte Carlo) used to generate the resulting probabilistic damage estimates. To this end, three distinct methods are demonstrated for rapid sampling that utilize surrogate modeling and exploit various degrees of parallelism for leveraging HPC. The accuracy and computational efficiency of the methods are compared on the problem of strain-based crack identification in thin plates. While each approach has inherent problem-specific strengths and weaknesses, all approaches are shown to provide accurate probabilistic damage diagnoses and several orders of magnitude computational speedup relative to a baseline Bayesian diagnosis implementation.

Warner, James E.

Electra: A Modular-Based Expansion of NASA's Supercomputing Capability

NASA has increasingly relied on high-performance computing (HPC) re- sources for computational modeling, simulation, and data analysis to meet the science and engineering goals of its missions in space exploration, aeronautics, and Earth and space science. The NASA Advanced Supercomputing (NAS) Division at Ames Research Center in Silicon Valley, Calif., hosts NASA’s premier supercomputing resources, integral to achieving and enhancing the success of the agency’s missions. NAS provides a balanced environment, funded under the High-End Computing Capability (HECC) project, comprised of world-class supercomputers, including its flagship distributed-memory cluster, Pleiades; high-speed networking; and massive data storage facilities, along with multi-disciplinary support teams for user support, code porting and optimization, and large-scale data analysis and scientific visualization. However, as scientists have increased the fidelity of their simulations and engineers are conducting larger parameter-space studies, the requirements for supercomputing resources have been growing by leaps and bounds. With the facility housing the HECC systems reaching its power and cooling capacity, NAS undertook a prototype project to investigate an alternative approach for housing supercomputers. Modular supercomputing, or container-based computing, is an innovative concept for expanding NASA’s HPC capabilities. With modular supercomputing, additional containers—similar to portable storage pods—can be connected together as needed to accommodate the agency’s ever-increasing demand for computing resources. In addition, taking advantage of the local weather permits the use of cooling technologies that would additionally save energy and reduce annual water usage. The first stage of NASA’s Modular Supercomputing Facility (MSF) prototype, which resulted in a 1,000 square-foot module on a concrete pad with room for 16 compute racks, was completed in Fall 2016 and an SGI (now HPE) computer system, named Electra, was deployed there in early 2017. Cooling is performed via an evaporative system built into the module, and preliminary experience shows a Power Usage Effectiveness (PUE) measurement of 1.03. Electra achieved over a petaflop on the LINPACK benchmark, sufficient to rank number 96 on the November 2016 TOP500 list [14]. The system consists of 1,152 InfiniBand-connected Intel Xeon Broadwell-based nodes. Its users access their files on a facility-wide file system shared by all HECC compute assets via Mellanox MetroX InfiniBand extenders, which connect the Electra fabric to Lustre routers in the primary facility over fiber-optic links about 900 feet long. The MSF prototype has exceeded expectations and is serving as a blueprint for future expansions. In the remainder of this chapter, we detail how modular data center technology can be used to expand an existing compute resource. We begin by describing NASA’s requirements for supercomputing and how resources were provided prior to the integration of the Electra module-based system.

Biswas, Rupak

Scaling Evaluation of Ice-Crystal Icing on a Modern Turbofan Engine in PSL Using the COMDES-MELT Code

This paper presents preliminary ice-crystal icing (ICI) altitude scaling evaluation results of a Honeywell Uncertified Research Engine (HURE) that was tested in the NASA Glenn Research Center Propulsion Systems Laboratory (PSL) during January of 2018. This engine geometry features a hidden core design to keep the core less exposed. The engine was fitted with internal video cameras to observe various ice buildup processes at multiple selected locations within the engine core flow path covering the fan stator, the splitter-lip/shroud/strut, and the high pressure compressor (HPC) variable inlet guide vane (IGV) regions. The potential ice accretion risk was pre-determined to occur by using NASA’s in-house 1D Engine Icing Risk assessment code, COMDES-MELT. The code was successful in predicting the risk of ice accretion in adiabatic regions like the fan-stator of the HURE at specific engine operating points. However at several operating points during the test, liquid water was observed running along the shroud toward the variable IGV of the HPC regions with an air temperature well below freezing, thus no particle melting could have occurred due to heating from the air alone. It was reasoned that other sources of heat were present in that region. To account for these heat sources the inlet total temperature was adjusted to give a wet bulb temperature of 24 F below the standard minimum wet bulb temperature of 492 R to allow ice to accrete in the splitter-lip/shroud/strut region, which was determined from a reference case where hard ice was observed in that region. With that adjustment the COMDES-MELT code was successful in providing operating points where there was a risk of ice accretion during the test campaign. In addition to calculating possible conditions at different selected lower altitudes, simulations were run to determine potential inlet conditions that could lead to ice-crystal accretion along the prescribed stations where the cameras were available. From there, scaled test conditions were determined by best matching the following three icing related parameters of the reference condition: (1) the local air total wet bulb temperature, (2) the local ice crystal cloud melt ratio and (3) the engine fan face ice/water to air mass flux ratio of the ice crystal cloud. Instantaneous images taken from the time-lapsed movies of ice buildup were used along with the relevant thermodynamic data of air, water vapor and local icing condition to help evaluate how closely the proposed altitude scaling method could be used in ground based test facility to duplicate selected reference ICI features observed at specific location inside this engine at different scale altitudes. Discussions on observed limitation for engine icing scaling application from this test campaign and needed improvement are provided. A scaling test procedure to help identify potential ICI risk conditions and possible ice accretion locations of a new turbofan engine is evaluated in PSL.

Engine Testing in PSL

Cloud Computing Methods for Near Rectilinear Halo Orbit Trajectory Design

Complicated mission design problems require innovative computational solutions. As spacecraft depart from a proposed Gateway in a Near Rectilinear Halo Orbit (NRHO), recontact analysis is required to avoid risk of collision and ensure safe operations. Escape dynamics from NRHOs are governed by multiple gravitational bodies, yielding a trajectory design space that is exhaustively large. This paper summarizes the recontact analysis for departure from the NRHO and describes how the Deep Space Trajectory Explorer (DSTE) trajectory design software incorporates high performance cloud computing to compute and visualize the orbit design space. Recent focus on exploration missions to cislunar space has kindled accelerated interest in multibody orbit solutions. Trajectory analysis in the presence of multiple gravity fields is complex, and innovative computational tools are needed to simplify complicated design spaces, to generate large quantities of data quickly, and to visualize the output for user accessibility. The Gateway mission is a prime example. The Gateway1 is proposed as a human outpost in deep space. The current baseline orbit for the Gateway is a Near Rectilinear Halo Orbit (NRHO) near the Moon.2 The NRHO exists in a regime that experiences the gravitational effects of the Earth and the Moon simultaneously, complicating orbit analysis. The mission design process benefits greatly from updated computational tools for multibody missions like the Gateway. As an example, consider the problem of assessing the risk of collision in an NRHO. As a staging location to missions to the lunar surface and beyond the Earth-Moon system, the Gateway will experience spacecraft and other objects regularly arriving and departing. Departing objects potentially include spent logistics modules, visiting crew vehicles, debris objects, wastewater particles, and cubesats. Each departure is governed by the dynamics of the Gateway orbit and the surrounding dynamical environment. Over time, any unmaintained object in such an orbit eventually departs due to the small instabilities associated with the NRHOs. A separation maneuver speeds the departure from the NRHO, but the effects of the maneuver on the spacecraft behavior depend on the location, magnitude, and direction of the burn. Escape dynamics from the NRHO with regard to these maneuver options open up an enormous potential trajectory design space where subtle changes in input can produce dramatically large changes in the results. Any departing object must avoid recontacting the Gateway as it leaves the lunar vicinity, and a recontact analysis thus involves a significant number of computations and extensive output data. To explore the dynamics of this extensive design space, the Deep Space Trajectory Explorer3 (DSTE) trajectory design software incorporates new High Performance Computing (HPC) services and novel interactive visualizations. This paper details the HPC and cloud infrastructure techniques that are implemented in the DSTE, applying the new capabilities to analysis of recontact risk with the Gateway in NRHO. NEAR RECTILINEAR HALO ORBITS The Gateway is planned to fly in a lunar NRHO as its baseline orbit. The NRHO families of orbits are subsets of the larger halo families, which originate from planar orbits near the L1 and L2 libration points; the Earth-Moon L2 halo family appears in Figure 1. Each halo orbit is perfectly periodic in the Circular Restricted 3-Body Problem (CR3BP) and becomes a quasi-periodic orbit in a higher fidelity ephemeris force model. The NRHOs are defined as those members of the halo family with bounded stability properties;2 they pass near the Moon at perilune and are nearly polar. Families exist with apolunes located both above the lunar north pole and above the lunar south pole; the Gateway is planned to reside in a southern L2 NRHO in a 9:2 resonance with the lunar synodic period. The 9:2 NRHO is characterized by a period of about 6.5 days, a perilune radius of about 3,500 km, and an apolune radius of about 71,000 km; it is strongly affected by the gravity of both the Earth and the Moon simultaneously. This NRHO offers extended communications with assets on the south pole of the Moon,4 as well as low-cost orbit maintenance and attitude control,5 favorable eclipse avoidance properties,6 and inexpensive transfers from Earth and to other destinations.5,7 The NRHO portion of the southern L2 halo family is highlighted in black in Figure 1, and the 9:2 NRHO appears in blue.

Phillips, Sean M.

Dynamic Analysis of the STARC-ABL Propulsion System

In the pursuit of Electrified Aircraft Propulsion (EAP), much of the attention is on the development of hybrid electric concept vehicles and their propulsion systems from a steady state performance perspective. While it is steady-state performance that largely determines the efficiency of civil air transports, engine operability and transient performance define constraints for the steady state design that impact efficiency and system viability. Neglecting dynamics and control technologies can result in an over-designed, sub optimal propulsion system or a concept that is not feasible. Thus, dynamic system studies were conducted on the propulsion system of the conceptual aircraft design known as the Single-aisle Turboelectric AiRCraft with Aft Boundary Layer propulsor (STARC-ABL). This paper describes the development of a controller to verify the baseline concept's feasibility from an operability perspective. Further, studies were conducted to identify excessive stability margin in the baseline design that could be traded for potential benefits in efficiency through an engine re design. This study revealed the potential to reduce the high pressure compressor (HPC) stall margin by 3%. Finally, a study was conducted to investigate the potential benefit of adding energy storage to the STARC-ABL concept that further improves operability and enables more gains in engine efficiency and performance. The energy storage provided an additional 0.5% stall margin can be removed from the HPC.

STARC-ABL

Dynamic Analysis of the STARC-ABL Propulsion System

In the pursuit of Electrified Aircraft Propulsion (EAP), much of the attention is on the development of hybrid electric concept vehicles and their propulsion systems from a steady state performance perspective. While it is steady-state performance that largely determines the efficiency of civil air transports, engine operability and transient performance define constraints for the steady state design that impact efficiency and system viability. Neglecting dynamics and control technologies can result in an over-designed, sub optimal propulsion system or a concept that is not feasible. Thus, dynamic system studies were conducted on the propulsion system of the conceptual aircraft design known as the Single-aisle Turboelectric AiRCraft with Aft Boundary Layer propulsor (STARC-ABL). This paper describes the development of a controller to verify the baseline concept's feasibility from an operability perspective. Further, studies were conducted to identify excessive stability margin in the baseline design that could be traded for potential benefits in efficiency through an engine re design. This study revealed the potential to reduce the high pressure compressor (HPC) stall margin by 3%. Finally, a study was conducted to investigate the potential benefit of adding energy storage to the STARC-ABL concept that further improves operability and enables more gains in engine efficiency and performance. The energy storage provided an additional 0.5% stall margin can be removed from the HPC.

STARC-ABL

Dynamic Analysis of the STARC-ABL Propulsion System

In the pursuit of Electrified Aircraft Propulsion (EAP), much of the attention is on the development of hybrid electric concept vehicles and their propulsion systems from a steady state performance perspective. While it is steady-state performance that largely determines the efficiency of civil air transports, engine operability and transient performance define constraints for the steady state design that impact efficiency and system viability. Neglecting dynamics and control technologies can result in an over-designed, sub optimal propulsion system or a concept that is not feasible. Thus, dynamic system studies were conducted on the propulsion system of the conceptual aircraft design known as the Single-aisle Turboelectric AiRCraft with Aft Boundary Layer propulsor (STARC-ABL). This paper describes the development of a controller to verify the baseline concept's feasibility from an operability perspective. Further, studies were conducted to identify excessive stability margin in the baseline design that could be traded for potential benefits in efficiency through an engine re design. This study revealed the potential to reduce the high pressure compressor (HPC) stall margin by 3%. Finally, a study was conducted to investigate the potential benefit of adding energy storage to the STARC-ABL concept that further improves operability and enables more gains in engine efficiency and performance. The energy storage provided an additional 0.5% stall margin can be removed from the HPC.

dynamic systems analysis

Using Big Data Technologies with Earth Science Data in HDF5: HDF5 Scalable Solutions

HDF5 (Hierarchical Data Format 5) is open-source, high-performance software that consists of an abstract data model, library, and fileformat used for storing and managing extremely large and/or complex data collections. NASA Earth Observing System (EOS) Data and Information Systems use HDF5 as an archival format to store remote sensing data from EOS satellites. HDF5 is also used to store other types of Geoscience and Strophysical data, e.g., seismic data and data from Low-Frequency Array (LOFAR) radio telescopes. Data stored in HDF5 has reached tens of petabytes and is growing at an accelerated rate.With the growing amout of HDF5 Earth Science data to analyze and process, scientists need to adopt big data technologies including new storage paradigms such as cloud and object storage. To run models and perform data analysis they also need to utilizied efficient and diverse ways to access data, from high-performance computing's (HPC) Message Passing Interface (MPI) I/O and deep memory hierarchies (DMH) to non-HPC frameworks such as Apache Hadoop, Spark, and Drill. The HDF Group continually works to enable usage of big data technologies in HDF software.

Knox, Larry

A High-Performance Computing Predictive GNSS Performance Monitor for Autonomous Air Vehicles in Urban Environments

This report offers analysis and design insights for leveraging High-Performance Computing (HPC) to predict line-of-sight (LOS) Global Navigation Satellite System (GNSS) availability in a city. This work is motivated by the emerging fields of Advanced and Urban Air Mobility (AAM/UAM), where regulatory authorities are seeking city-scale, meter-resolution risk forecasting in order to safely integrate new flight missions with existing urban life and infrastructure. This work addresses the technical challenge of efficiently computing urban GNSS satellite visibility to predict GNSS performance metrics under these requirements. We present a new HPC-optimized shadow casting algorithm variant as a ray-based approach to forecasting satellite visibility. We apply this algorithm variant in a software-defined prognostic service which generates a GNSS navigation risk-correlated map as a path planning-style potential field. We detail dominant computational burdens, viable simplifying assumptions, and different algorithmic implementations, intending to demonstrate a baseline of computation time needed by each stage in such a service. We conclude by analyzing the prototype service’s prediction accuracy compared to receiver data from Corpus Christi, Texas. This informs design trade-offs along the dimensions of hardware, computation time, and tolerable forecasting error (including proportions of false positives and false negatives).

GNSS

NASA Center for Climate Simulation (NCCS) Advanced Technology AT5 Virtualized Infiniband Report

The NCCS is part of the Computational and Information Sciences and Technology Office (CISTO) of Goddard Space Flight Center's (GSFC) Sciences and Exploration Directorate. The NCCS's mission is to enable scientists to increase their understanding of the Earth, the solar system, and the universe by supplying state-of-the-art high performance computing (HPC) solutions. To accomplish this mission, the NCCS (https://www.nccs.nasa.gov) provides high performance compute engines, mass storage, and network solutions to meet the specialized needs of the Earth and space science user communities

cloud

GPU Implementation of the OVERFLOW CFD Code

The high-performance computing (HPC) landscape is quickly changing to systems where most of the performance comes from specialized chips, specifically graphics processing units (GPUs). Such GPU systems are throughput machines, where efficient use of the GPU often requires code refactoring to expose a few orders of magnitude more fine grain parallelism than was previously used on the CPU. Recent modifications to OVERFLOW, an overset, structured grid, computational fluid dynamics flow solver, written in Fortran will be presented. These modifications include both code modernization efforts and algorithmic changes to enable OVERFLOW to efficiently utilize GPUs. Many of these algorithmic changes would likely also be applicable for other structured grid, stencil-based codes wanting to utilize GPUs. The capabilities that have been ported to run on the GPUs are presented, along with the performance gains of the GPU version relative the CPU version of OVERFLOW.

GPU Programming

GPU Implementation of the OVERFLOW CFD Code

The high-performance computing (HPC) landscape is quickly changing to systems where most of the performance comes from specialized chips, specifically graphics processing units (GPUs). Such GPU systems are throughput machines, where efficient use of the GPU often requires code refactoring to expose a few orders of magnitude more fine grain parallelism than was previously used on the CPU. Recent modifications to OVERFLOW, an overset, structured grid, computational fluid dynamics flow solver, written in Fortran will be presented. These modifications include both code modernization efforts and algorithmic changes to enable OVERFLOW to efficiently utilize GPUs. Many of these algorithmic changes would likely also be applicable for other structured grid, stencil-based codes wanting to utilize GPUs. The capabilities that have been ported to run on the GPUs are presented, along with the performance gains of the GPU version relative the CPU version of OVERFLOW.

GPU Programming

Implementation of BT, SP, LU, and FT of NAS Parallel Benchmarks in Java

A number of Java features make it an attractive but a debatable choice for High Performance Computing. We have implemented benchmarks working on single structured grid BT,SP,LU and FT in Java. The performance and scalability of the Java code shows that a significant improvement in Java compiler technology and in Java thread implementation are necessary for Java to compete with Fortran in HPC applications.

Schultz, Matthew