Engineering PapersSearch

SEARCH · Engineering Papers

Results for “processor performance factors”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

120 records · Page 7

Algorithms and Application of Sparse Matrix Assembly and Equation Solvers for Aeroacoustics

An algorithm for symmetric sparse equation solutions on an unstructured grid is described. Efficient, sequential sparse algorithms for degree-of-freedom reordering, supernodes, symbolic/numerical factorization, and forward backward solution phases are reviewed. Three sparse algorithms for the generation and assembly of symmetric systems of matrix equations are presented. The accuracy and numerical performance of the sequential version of the sparse algorithms are evaluated over the frequency range of interest in a three-dimensional aeroacoustics application. Results show that the solver solutions are accurate using a discretization of 12 points per wavelength. Results also show that the first assembly algorithm is impractical for high-frequency noise calculations. The second and third assembly algorithms have nearly equal performance at low values of source frequencies, but at higher values of source frequencies the third algorithm saves CPU time and RAM. The CPU time and the RAM required by the second and third assembly algorithms are two orders of magnitude smaller than that required by the sparse equation solver. A sequential version of these sparse algorithms can, therefore, be conveniently incorporated into a substructuring for domain decomposition formulation to achieve parallel computation, where different substructures are handles by different parallel processors.

Watson, W. R.

Electromagnetic Radial Forces in a Hybrid Eight-Stator-Pole, Six-Rotor-Pole Bearingless Switched-Reluctance Motor

Analysis and experimental measurement of the electromagnet force loads on the hybrid rotor in a novel bearingless switched-reluctance motor (BSRM) have been performed. A BSRM has the combined characteristics of a switched-reluctance motor and a magnetic bearing. The BSRM has an eight-pole stator and a six-pole hybrid rotor, which is composed of circular and scalloped lamination segments. The hybrid rotor is levitated using only one set of stator poles. A second set of stator poles imparts torque to the scalloped portion of the rotor, which is driven in a traditional switched reluctance manner by a processor. Analysis was done for nonrotating rotor poles that were oriented to achieve maximum and minimum radial force loads on the rotor. The objective is to assess whether simple one-dimensional magnetic circuit analysis is sufficient for preliminary evaluation of this machine, which may exhibit strong three-dimensional electromagnetic field behavior. Two magnetic circuit geometries, approximating the complex topology of the magnetic fields in and around the hybrid rotor, were employed in formulating the electromagnetic radial force equations. Reasonable agreement between the experimental results and the theoretical predictions was obtained with typical magnetic bearing derating factors applied to the predictions.

Morrison, Carlos R.

Thermal insulation testing method and apparatus

A test apparatus and method of its use for evaluating various performance aspects of a test specimen is disclosed. A chamber within a housing contains a cold mass tank with a contact surface in contact with a first surface of a test specimen. The first surface of the test specimen is spaced from the second surface of the test specimen by a thickness. The second surface of the test specimen is maintained at a desired warm temperature. The first surface is maintained at a constant temperature by a liquid disposed within the cold mass tank. A boil-off flow rate of the gas is monitored and provided to a processor along with the temperature of the first and second surfaces of the test specimen. The processor calculates thermal insulation values of the test specimen including comparative values for heat flux and apparent thermal conductivity (k-value). The test specimen may be placed in any vacuum pressure level ranging from about 0.01 millitorr to 1,000,000 millitorr with different residual gases as desired. The test specimen may be placed under a mechanical load with the cold mass tank and another factors may be imposed upon the test specimen so as to simulate the actual use conditions.

Fesmire, James E.

Multiprocessing on supercomputers for computational aerodynamics

Very little use is made of multiple processors available on current supercomputers (computers with a theoretical peak performance capability equal to 100 MFLOPs or more) in computational aerodynamics to significantly improve turnaround time. The productivity of a computer user is directly related to this turnaround time. In a time-sharing environment, the improvement in this speed is achieved when multiple processors are used efficiently to execute an algorithm. The concept of multiple instructions and multiple data (MIMD) through multi-tasking is applied via a strategy which requires relatively minor modifications to an existing code for a single processor. Essentially, this approach maps the available memory to multiple processors, exploiting the C-FORTRAN-Unix interface. The existing single processor code is mapped without the need for developing a new algorithm. The procedure for building a code utilizing this approach is automated with the Unix stream editor. As a demonstration of this approach, a Multiple Processor Multiple Grid (MPMG) code is developed. It is capable of using nine processors, and can be easily extended to a larger number of processors. This code solves the three-dimensional, Reynolds averaged, thin-layer and slender-layer Navier-Stokes equations with an implicit, approximately factored and diagonalized method. The solver is applied to generic oblique-wing aircraft problem on a four processor Cray-2 computer. A tricubic interpolation scheme is developed to increase the accuracy of coupling of overlapped grids. For the oblique-wing aircraft problem, a speedup of two in elapsed (turnaround) time is observed in a saturated time-sharing environment.

Yarrow, Maurice

Electromagnetic Forces in a Hybrid Magnetic-Bearing Switched-Reluctance Motor

Analysis and experimental measurement of the electromagnetic force loads on the hybrid rotor in a novel hybrid magnetic-bearing switched-reluctance motor (MBSRM) have been performed. A MBSRM has the combined characteristics of a switched-reluctance motor and a magnetic bearing. The MBSRM discussed in this report has an eight-pole stator and a six-pole hybrid rotor, which is composed of circular and scalloped lamination segments. The hybrid rotor is levitated using only one set of four stator poles, while a second set of four stator poles imparts torque to the scalloped portion of the rotor, which is driven in a traditional switched reluctance manner by a processor. Static torque and radial force analysis were done for rotor poles that were oriented to achieve maximum and minimum radial force loads on the rotor. The objective is to assess whether simple one-dimensional magnetic circuit analysis is sufficient for preliminary evaluation of this machine, which may exhibit strong three-dimensional electromagnetic field behavior. Two magnetic circuit geometries, approximating the complex topology of the magnetic fields in and around the hybrid rotor, were employed in formulating the electromagnetic radial force equations. Reasonable agreement between the experimental and the theoretical radial force loads predictions was obtained with typical magnetic bearing derating factors applied to the predictions.

Morrison, Carlos R.

RadSTAR L-Band Imaging Scatterometer: Performance Assessment

RadSTAR is an instrument development program aimed at combining a radiometer and a scatterometer system into a highly compact configuration that uses a single, electronically scanned antenna to provide co-located and simultaneous measurements of emission and backscatter for airborne and spaceborne applications [I]. The program was designed to map soil moisture and ocean salinity, both important components of the water cycle, and to map sea ice density and thickness, an important factor in ocean-atmosphere heat exchange in Polar Regions. The accuracy in estimation of these and a number of other Earth science parameters can be greatly enhanced by providing the co-aligned radar/radiometer microwave measurements. For instance, radiometer estimates of soil moisture from soil emission are affected by emission from vegetation, and from the roughness of the surface. Complementary measurements using the scatterometer can be used to evaluate the vegetation and surface roughness effects. Hence, the combined observations can provide an improved estimate. As with soil moisture, the ocean salinity is a function of the microwave emission from the sea surface temperature (SST) and sea roughness. There, the addition of radar backscatter measurements of sea roughness enables the correction of the emissivity and provide more accurate estimates of ocean salinity. Similar arguments can be made for other important Earth science parameters. This paper discusses the RadSTAR program, the radar system design, calibration, and digital beamforming techniques, and presents preliminary analysis of the data collected during the test flights. The data sets obtained during the flights and during the radar calibration in the anechoic chamber are also employed to asses the performance of the radar. The paper also discusses the Digital Beamforming Synthetic Aperture Radar (DBSAR) processor, a real-time processor recently developed for the LIS instrument which enables beam synthesis, fine resolutions, and large swaths.

Rincon, Rafael

A Tool for Automatic Data Distribution for CFD Applications on Structured Grids

Development of HPF versions of NPB and ARC3D has shown that HPF provides an efficient, concise way to express parallelism and to organize data traffic. The use of HPF, as noted in the papers, requires an intimate knowledge of the applications and a detailed analysis of data affinity, data movement, and data granularity. To simplify and accelerate the task of developing HPF versions of existing CFD applications we have designed and implemented ADAPT (Automatic Data Alignment and Placement Tool). ADAPT analyzes a CFD application working on a single structured grid and generates HPF TEMPLATE, (RE)DISTRIBUTION, ALIGNMENT, and INDEPENDENT directives. The directives can be generated on the nest level, subroutine level, application level, or on the application interface level. ADAPT annotates an existing CFD FORTRAN application, performing computations on single or multiple grids. On each grid the application is considered as a sequence of operators, each applied to a set of variables defined in a particular grid domain. ADAPT automatically detects implicit operators (i.e., having data dependences) and explicit operators (without data dependences). For parallelization of an explicit operator ADAPT creates a template for the operator domain, aligns arrays used in the operator with the template, distributes the template, and declares the loops over the distributed dimensions as INDEPENDENT. For parallelization of an implicit operator, the distribution of the operator's domain should be consistent with the operator's dependences. Any dependence between sections distributed on different processors would preclude parallelization if the compiler does not have an ability to pipeline computations. If a data distribution is "orthogonal" to the dependences of an implicit operator, then the loop which implements the operator can be declared as INDEPENDENT. ADAPT starts with an analysis of array index expressions of the loop nests. For each pair of arrays referenced in an assignment statement, it generates an arc in the alignment graph and annotates it with an affinity relation. The template, alignment, and distribution directives for a particular loop nest are then derived from a transitive closure of the affinity relation. A compromise of data distributions in different nests and subroutines is achieved by merging annotated alignment graphs for adjacent nests/stibroutine calls in the nest/call graph of the application in the process called distribution lifting. ADAPT has been implemented as a C++ program running in conjunction with a parallelization tool called CAPTools. ADAPT uses the parse tree, interprocedural analysis and application database generated by CAPTools. It also uses the Directed Graph class, initially implemented in p2d2 (parallel debugger oi distributed programs), and some other classes supporting symbolic computations. ADAPT uses data distribution techniques described. ADAPT was tested with ARC3D and the FT benchmark and has demonstrated a code performance within a factor of 1.5 of handwritten versions.

Frumkin, Michael

Toward Real Time Neural Net Flight Controllers

NASA Ames Research Center has an ongoing program in neural network control technology targeted toward real time flight demonstrations using a modified F-15 which permits direct inner loop control of actuators, rapid switching between alternative control designs, and substitutable processors. An important part of this program is the ACTIVE flight project which is examining the feasibility of using neural networks in the design, control, and system identification of new aircraft prototypes. This paper discusses two research applications initiated with this objective in mind: utilization of neural networks for wind tunnel aircraft model identification and rapid learning algorithms for on line reconfiguration and control. The first application involves the identification of aerodynamic flight characteristics from analysis of wind tunnel test data. This identification is important in the early stages of aircraft design because complete specification of control architecture's may not be possible even though concept models at varying scales are available for aerodynamic wind tunnel testing. Testing of this type is often a long and expensive process involving measurement of aircraft lift, drag, and moment of inertia at varying angles of attack and control surface configurations. This information in turn can be used in the design of the flight control systems by applying the derived lookup tables to generate piece wise linearized controllers. Thus, reduced costs in tunnel test times and the rapid transfer of wind tunnel insights into prototype controllers becomes an important factor in more efficient generation and testing of new flight systems. NASA Ames Research Center is successfully applying modular neural networks as one way of anticipating small scale aircraft model performances prior to testing, thus reducing the number of in tunnel test hours and potentially, the number of intermediate scaled models required for estimation of surface flow effects.

Jorgensen, C. C.

A Parallelized Oxidation-Driven Surface Recession Framework in DSMC Code, SPARTA

Spacecrafts rely on ablative thermal protection systems (TPS) made of composites consisting of a carbon-based reinforcement and a polymeric matrix. These materials are designed to withstand high-temperature oxidation and surface recession during re-entry into the Earth's atmosphere. However, ablation occurs due to a complex interplay of thermal, mechanical, and chemical factors, making it challenging to determine the individual impact of each on the TPS's overall degradation. In this study, we have developed an ablation model that can leverage a finite rate carbon oxidation model to predict material recession and surface states more accurately. Stochastic PArallel Rarified-gas Time-accurate Analyzer (SPARTA), a direct-simulation Monte Carlo (DSMC) code, is modified to allow oxidation-driven ablation of implicitly defined carbon surfaces. In SPARTA, implicit surfaces are generated from the grid corner point values via a marching cubes algorithm, therefore creating a new set of surface elements every time ablation is performed. The finite-rate oxidation model developed by Gopalan et. al can perform both gas-surface and pure-surface reactions and is now adapted to tally surface data on a per grid cell basis. The ablation functionality was also adjusted so once the reactions have occurred, the number of reactions leading to CO formation can be converted to corner point reduction values; therefore, carbon removal is directly proportional to surface recession. We also briefly discuss some unique challenges associated with parallelizing this dynamic surface state and geometry. Finally, we analyze the performance of this parallelized implicit chemistry model with simple 2D and 3D benchmark cases by producing surface state statistics, area changes over time, and visualization across a range of surface temperatures and processors with and without load-balancing.

DSMC

Forecasting of Storm-Surge Floods Using ADCIRC and Optimized DEMs

Increasing the accuracy of storm-surge flood forecasts is essential for improving preparedness for hurricanes and other severe storms and, in particular, for optimizing evacuation scenarios. An interactive database, developed by WorldWinds, Inc., contains atlases of storm-surge flood levels for the Louisiana/Mississippi gulf coast region. These atlases were developed to improve forecasting of flooding along the coastline and estuaries and in adjacent inland areas. Storm-surge heights depend on a complex interaction of several factors, including: storm size, central minimum pressure, forward speed of motion, bottom topography near the point of landfall, astronomical tides, and, most importantly, maximum wind speed. The information in the atlases was generated in over 100 computational simulations, partly by use of a parallel-processing version of the ADvanced CIRCulation (ADCIRC) model. ADCIRC is a nonlinear computational model of hydrodynamics, developed by the U.S. Army Corps of Engineers and the US Navy, as a family of two- and three-dimensional finite-element-based codes. It affords a capability for simulating tidal circulation and storm-surge propagation over very large computational domains, while simultaneously providing high-resolution output in areas of complex shoreline and bathymetry. The ADCIRC finite-element grid for this project covered the Gulf of Mexico and contiguous basins, extending into the deep Atlantic Ocean with progressively higher resolution approaching the study area. The advantage of using ADCIRC over other storm-surge models, such as SLOSH, is that input conditions can include all or part of wind stress, tides, wave stress, and river discharge, which serve to make the model output more accurate. To keep the computational load manageable, this work was conducted using only the wind stress, calculated by using historical data from Hurricane Camille, as the input condition for the model. Hurricane storm-surge simulations were performed on an eight-node Linux computer cluster. Each node contained dual 2-GHz processors, 2GB of memory, and a 40GB hard drive. The digital elevation model (DEM) for this region was specified using a combination of Navy data (over water), NOAA data (for the coastline), and optimized Interferometric Synthetic Aperture Radar data (over land). This high-resolution topographical data of the Mississippi coastal region provided the ADCIRC model with improved input with which to calculate improved storm-surge forecasts.

Valenti, Elizabeth

Ion Exchange Technology Development in Support of the Urine Processor Assembly

The urine processor assembly (UPA) on the International Space Station (ISS) recovers water from urine via a vacuum distillation process. The distillation occurs in a rotating distillation assembly (DA) where the urine is heated and subjected to sub-ambient pressure. As water is removed, the original organics, salts, and minerals in the urine become more concentrated and result in urine brine. Eventually, water removal will concentrate the urine brine to super saturation of individual constituents, and precipitation occurs. Under typical UPA DA operating conditions, calcium sulfate or gypsum is the first chemical to precipitate in substantial quantity. During preflight testing with ground urine, the UPA achieved 85% water recovery without precipitation. However, on ISS, it is possible that crewmember urine can be significantly more concentrated relative to urine from ground donors. As a result, gypsum precipitated in the DA when operating at water recovery rates at or near 85%, causing the failure and subsequent re14 NASA Tech Briefs, September 2013 placement of the DA. Later investigations have demonstrated that an excess of calcium and sulfate will cause precipitation at water recovery rates greater than 70%. The source of the excess calcium is likely physiological in nature, via crewmembers' bone loss, while the excess sulfate is primarily due to the sulfuric acid component of the urine pretreatment. To prevent gypsum precipitation in the UPA, the Precipitation Prevention Project (PPP) team has focused on removing the calcium ion from pretreated urine, using ion exchange resins as calcium removal agents. The selectivity and effectiveness of ion exchange resins are determined by such factors as the mobility of the liquid phase through the polymer matrix, the density of functional groups, type of functional groups bound to the matrix, and the chemical characteristics of the liquid phase (pH, oxidation potential, and ionic strength). Previous experience with ion exchange resins has demonstrated that the most effective implementation for an ion exchange resin is a cartridge, or column, in which the resin is contained. Based on the results of equilibrium and sub-scale dynamic column testing, a possible solution for mitigating the calcium precipitation issue on the ISS has been identified. From an original pool of 13 ion exchange resins, two candidates have been identified that demonstrate substantial calcium removal on the sub-scale. The dramatic reduction in resin performance from published calcium uptake demonstrates the need for thorough evaluation of resins at the low pH and strong oxidizing environment present in the UPA. Chemical variations in the influent (calcium concentrations and pretreatment dosing) appear to have a noticeable impact on the calcium capacity of the resin. Low calcium concentrations and high pretreatment dosing will likely result in a decrease in calcium capacity. Conversely, low pre trea t - ment dosing will likely result in an increase in calcium capacity. In contrast, investigations at a variety of flow rates, length-to-diameter ratios, resin volumes, and flow regimes (continuous versus pulsed) show that changes in physical parameters do not have substantial impacts on resin performance in the very low specific velocity ranges of interest. This result is particularly useful because most commercial applications at higher specific velocities do show a relatively strong relationship between flow and capacity. The lack of a strong relationship will allow more flexibility in the implementation of an ion exchange bed for flight. Verification of subscale tests with flight-scale resin beds is recommended prior to implementation in the on-orbit UPA.

Mitchell, Julie

Dynamics and Control of a Disordered System in Space

In this paper, we present some ideas regarding the modeling, dynamics and control aspects of granular spacecraft. Granular spacecraft are complex multibody systems composed of a spatially disordered distribution of a large number of elements, for instance a cloud of N grains in orbit, with N greater than 10(exp 3). These grains can be large (Cubesat-size) or small (mm-size), and can be active, i.e., a fully equipped vehicle capable sensing their own position and attitude, and enabled with propulsion means, or entirely passive. The ultimate objective would be to study the behavior of the single grains and of large ensembles of grains in orbit and to identify ways to guide and control the shape of a cloud composed of these grains so that it can perform a useful function in space, for instance, as an element of an optical imaging system for astrophysical applications. This concept, in which the aperture does not need to be continuous and monolithic, would increase the aperture size several times compared to large NASA observatories such as ATLAST, allowing for a true Terrestrial Planet Imager that would be able to resolve exo-planet details and do meaningful spectroscopy on distant world. In the paper, we address the modeling and autonomous operation of a distributed assembly (the cloud) of large numbers of highly miniaturized space-borne elements (the grains). A multi-scale, multi-physics model is proposed of the dynamics of the cloud in orbit, as well as a control law for cloud shape maintenance, and preliminary simulation studies yield an estimate of the computational effort, indicating a scale factor of approximately N(exp 1.4) as a function of the number of grains. A granular spacecraft can be defined as a collection of a large number of space-borne elements (in the 1000s) designed and controlled such that a desirable collective behavior emerges, either from the interactions among neighboring grains, and/or between the grains and the environment. In this paper, each grain is considered to be a highly miniaturized spacecraft which has limited size and mass, hence it has limited actuation, limited propulsive capability, limited power, limited sensing, limited communication, limited computational resources, limited range of motion, limited lifetime, and may be expendable. The modeling and dynamics of clouds of vehicles is more challenging than with conventional vehicles because we are faced with a probabilistic vehicle composed of a large number of physically disconnected vehicles. First, different scales of motion occur simultaneously in a cloud: translations and rotations of the cloud as a whole (macro-dynamics), relative rotation and translation of one cloud member with respect to another (meso-dynamics), and individual cloud member dynamics (micro-dynamics). Second, the control design needs to be tolerant of the system complexity, of the system architecture (centralized vs. decentralized large scale system control) as well as robust to un-modeled dynamics and noise sources. Figure 1, top left, shows the kinematic parameters of a 1000 element cloud in orbit. The motion of the system is described with respect to a local vertical-local horizontal (LV-LH) orbiting reference frame (x,y,z)=F(sub ORF) of origin O(sub ORF) which rotates with mean motion omega and orbital semi-major axis R(sub 0). The orbital geometry at the initial time is defined in terms of its six orbital elements, and the orbital dynamics equation for point O(sub ORF) is propagated forward in time under the influence of the gravitational field of the primary and other external perturbations, described below. The origin of this frame coincides with the initial position of the center of mass of the system, and the coordinate axes are z along the local vertical, x toward the flight direction, and y in the orbit normal direction. The assumptions we used to model the dynamics are as follows: 1) The inertial frame is fixed at Earth's center. 2) The orbiting Frame ORF follows Keplerian orbit. 3) the cloud system dynamics is referred to ORF. 4) the attitude of each grain uses the principal body frame as body fixed frame. 5) the atmosphere is assumed to be rigidly rotating with the Earth. Regarding the grains forming the cloud: 1) each grain is modeled as a rigid body; 2) a simple attitude estimator provides attitude estimates, 3) a simple guidance logic commands the position and attitude of each grain, 4) a simple local feedback controller based on PD control of local states is used to stabilize the attitude of the vehicle. Regarding the cloud: 1) the cloud as a whole is modeled as an equivalent rigid body in orbit, and 2) an associated graph establishes agent connectivity and enables coupling between modes of motion at the micro and macro scales; 3) a simple guidance and estimation logic is modeled to estimate and command the attitude of this equivalent rigid body; 4) a cloud shape maintenance controller is based on the dynamics of a stable virtual truss in the orbiting frame. Regarding the environmental perturbations acting on the cloud: 1) a non-spherical gravity field including JO (Earth's spherical field) zonal component, J2 (Earth's oblateness) and J3 zonal components is implemented; 2) atmospheric drag is modeled with an exponential model; 3) solar pressure is modeled assuming the Sun is inertially fixed; and 4) the Earth's magnetic field is model using an equivalent dipole model. The equations of motion are written in a referential system with respect to the origin of the orbiting frame and the state is propagated forward in time using an incremental predictor-corrector scheme. A representative cloud with varying number of grains is simulated to identify the limitations in computation time as the number of grains grows. We derive a control law to track a desired surface in the ORF (equivalently to maintain a reference cloud shape) by defining an error from a desired surface shape, and designing a control law that is exponentially stable and reduces the tracking error to zero. Figure 1 (top right) shows a comparison of various requirements for simulation of single spacecraft vs. granular spacecraft, indicating the high degree of complexity that needs to be taken into consideration. The ORF components of control force required by one of the grains is, for this particular case, in the micro-Newton range. However, no attempt has been made yet to reconfigure (or re-orient) the cloud configuration internally, for which forces in the milli-Newton level are expected, depending on the time required to do the reconfiguration. Figure 1, bottom, shows the computation time as a function of the number of grains, indicating an order N(exp 1.43) scaling on a 8 Gb, 1067 MHz RAM MacOSX computer with a 3.06 GHz Intel Core 2 Duo processor. With this metric, the same simulation for a system of N=1000 grains would take 5.4 hours, and 146 hours (i.e., 6 days) for a system with N=10,000 grains. Therefore, efficient ways to simulate this complex system, where not only the time scales of natural system dynamics, but also the sampling times of the Guidance, Navigation, and Control are included, remain to be explored. Additional details on the cloud modeling, dynamics, and control will be described in the paper.

simulation