Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “SMP”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Use Computer-Aided Tools to Parallelize Large CFD Applications

Porting applications to high performance parallel computers is always a challenging task. It is time consuming and costly. With rapid progressing in hardware architectures and increasing complexity of real applications in recent years, the problem becomes even more sever. Today, scalability and high performance are mostly involving handwritten parallel programs using message-passing libraries (e.g. MPI). However, this process is very difficult and often error-prone. The recent reemergence of shared memory parallel (SMP) architectures, such as the cache coherent Non-Uniform Memory Access (ccNUMA) architecture used in the SGI Origin 2000, show good prospects for scaling beyond hundreds of processors. Programming on an SMP is simplified by working in a globally accessible address space. The user can supply compiler directives, such as OpenMP, to parallelize the code. As an industry standard for portable implementation of parallel programs for SMPs, OpenMP is a set of compiler directives and callable runtime library routines that extend Fortran, C and C++ to express shared memory parallelism. It promises an incremental path for parallel conversion of existing software, as well as scalability and performance for a complete rewrite or an entirely new development. Perhaps the main disadvantage of programming with directives is that inserted directives may not necessarily enhance performance. In the worst cases, it can create erroneous results. While vendors have provided tools to perform error-checking and profiling, automation in directive insertion is very limited and often failed on large programs, primarily due to the lack of a thorough enough data dependence analysis. To overcome the deficiency, we have developed a toolkit, CAPO, to automatically insert OpenMP directives in Fortran programs and apply certain degrees of optimization. CAPO is aimed at taking advantage of detailed inter-procedural dependence analysis provided by CAPTools, developed by the University of Greenwich, to reduce potential errors made by users. Earlier tests on NAS Benchmarks and ARC3D have demonstrated good success of this tool. In this study, we have applied CAPO to parallelize three large applications in the area of computational fluid dynamics (CFD): OVERFLOW, TLNS3D and INS3D. These codes are widely used for solving Navier-Stokes equations with complicated boundary conditions and turbulence model in multiple zones. Each one comprises of from 50K to 1,00k lines of FORTRAN77. As an example, CAPO took 77 hours to complete the data dependence analysis of OVERFLOW on a workstation (SGI, 175MHz, R10K processor). A fair amount of effort was spent on correcting false dependencies due to lack of necessary knowledge during the analysis. Even so, CAPO provides an easy way for user to interact with the parallelization process. The OpenMP version was generated within a day after the analysis was completed. Due to sequential algorithms involved, code sections in TLNS3D and INS3D need to be restructured by hand to produce more efficient parallel codes. An included figure shows preliminary test results of the generated OVERFLOW with several test cases in single zone. The MPI data points for the small test case were taken from a handcoded MPI version. As we can see, CAPO's version has achieved 18 fold speed up on 32 nodes of the SGI O2K. For the small test case, it outperformed the MPI version. These results are very encouraging, but further work is needed. For example, although CAPO attempts to place directives on the outer- most parallel loops in an interprocedural framework, it does not insert directives based on the best manual strategy. In particular, it lacks the support of parallelization at the multi-zone level. Future work will emphasize on the development of methodology to work in a multi-zone level and with a hybrid approach. Development of tools to perform more complicated code transformation is also needed.

Jin, H.↗

Central Stars of Planetary Nebulae in the SMC

In FUSE cycle 3's program C056 we studied four Central Stars of Planetary Nebulae (CSPN) in the Small Magellanic Could. All FUSE observations have been successfully completed and have been reduced and analyzed. The observation of one object (SMP SMC 5) appeared to be off-target and no useful stellar flux was gathered. For another observation (SMP SMC 1) the voltage problems resulted in the loss of data from one of the SiC detectors, but we were still able to analyze the remaining data. The analysis and the results are summarized below. The FUSE data were reduced using the latest available version of the FUSE calibration pipeline (CALFUSE v2.4). The flux of these SMC post-AGB objects is at the threshold of FUSE S sensitivity, and the targets required many orbit-long exposures, each of which typically had low (target) count-rates. The background subtraction required special care during the reduction, and was done in a similar manner to our FUSE cycle 2 BOO1 objects. The resulting calibrated data from the different channels were compared in the overlapping regions for consistency. The final combined, extracted spectra of each target was then modeled to determine the stellar and nebular parameters. The FUSE spectra, combined with archival HST spectra, have been analyzed using stellar atmospheres codes such as TLUSTY and CMFGEN to derive photospheric and wind parameters of the central stars, and with ISM models to determine the amount and temperature of the surrounding atomic and molecular hydrogen. We have combined these results with those of our cycle 4 (D034) program (CSPN of the LMC) in Herald & Bianchi 2004a (paper in preparation, will be submitted to ApJ in June 2004). Two of the three SMC objects analyzed were found to have significantly lower stellar temperatures than had been predicted using nebular photoionization models, indicating either a hotter ionizing companion or the existence of strong shocks in the nebular environment. The analysis also revealed that some objects are surrounded by significant quantities of hot (e.g., 1000- 2000 K) molecular H2, similar to what we found for some LMC and Galactic CSPN (Herald & Bianchi 2002, 2003a, b, 2004b, c, d).

Bianchi, Luciana↗

Lightweight, Self-Deployable Wheels

Ultra-lightweight, self-deployable wheels made of polymer foams have been demonstrated. These wheels are an addition to the roster of cold hibernated elastic memory (CHEM) structural applications. Intended originally for use on nanorovers (very small planetary-exploration robotic vehicles), CHEM wheels could also be used for many commercial applications, such as in toys. The CHEM concept was reported in "Cold Hibernated Elastic Memory (CHEM) Expandable Structures" (NPO-20394), NASA Tech Briefs, Vol. 23, No. 2 (February 1999), page 56. To recapitulate: A CHEM structure is fabricated from a shape-memory polymer (SMP) foam. The structure is compressed to a very small volume while in its rubbery state above its glass-transition temperature (Tg). Once compressed, the structure can be cooled below Tg to its glassy state. As long as the temperature remains <Tg the structure remains compacted (in a cold hibernated state), even when the external compressive forces are removed. When the structure is subsequently heated above Tg, it returns to the rubbery state, in which a combination of elasticity and the SMP effect cause it to expand (deploy) to its original size and shape. Once thus deployed, the CHEM structure can be rigidified by cooling below Tg to the glassy state. The structure could be subsequently reheated above Tg and recompacted. The compaction/deployment/rigidification cycle could be repeated as many times as needed.

Chmielewski, Artur↗

Computer-Aided Parallelizer and Optimizer

The Computer-Aided Parallelizer and Optimizer (CAPO) automates the insertion of compiler directives (see figure) to facilitate parallel processing on Shared Memory Parallel (SMP) machines. While CAPO currently is integrated seamlessly into CAPTools (developed at the University of Greenwich, now marketed as ParaWise), CAPO was independently developed at Ames Research Center as one of the components for the Legacy Code Modernization (LCM) project. The current version takes serial FORTRAN programs, performs interprocedural data dependence analysis, and generates OpenMP directives. Due to the widely supported OpenMP standard, the generated OpenMP codes have the potential to run on a wide range of SMP machines. CAPO relies on accurate interprocedural data dependence information currently provided by CAPTools. Compiler directives are generated through identification of parallel loops in the outermost level, construction of parallel regions around parallel loops and optimization of parallel regions, and insertion of directives with automatic identification of private, reduction, induction, and shared variables. Attempts also have been made to identify potential pipeline parallelism (implemented with point-to-point synchronization). Although directives are generated automatically, user interaction with the tool is still important for producing good parallel codes. A comprehensive graphical user interface is included for users to interact with the parallelization process.

Jin, Haoqiang↗

Enabling Simulation Interoperability between International Standards in the Space Domain

Today, the design and development of space systems are conducted cooperatively by historical space agencies in this area such as NASA, ESA, Roscosmos, and JAXA together with their industrial partners. The space system lifecycle is characterized by high costs, uncertain conditions, and dangerous scenarios. To mitigate these issues, space agencies rely heavily on Modelling and Simulation as a key technology to support the analysis, design, and operation of space systems. To support large-scale distributed simulations, the scientific community has developed several standards to support the reuse and interoperability of simulation models such as IEEE 1516 High-Level Architecture (HLA), Real-time Platform Reference Federation Object Model (RPR FOM), Simulation Model Portability (SMP), and the novel Space Reference Federation Object Model (SpaceFOM). While the SpaceFOM standard has been specifically conceptualized for handling space systems, the other ones are more general-purpose and can be used to design and simulate generic complex systems. As a consequence, there is a lake of rules and guidelines to enable interoperability among these standards. The paper presents solutions and experiences for enabling interoperability and transferability of HLA, RPR FOM, and SMP simulation models with SpaceFOM.

Simulation↗

Calculation of transonic steady and oscillatory pressures on a low aspect ratio model and comparison with experiment

Pressure data measured by the British Royal Aircraft Establishment for the AGARD SMP tailplane are compared with results calculated using the transonic small perturbation code XTRAN3S. A brief description of the analysis is given and a recently developed finite difference grid is described. Results are presented for five steady and nine harmonically oscillating cases near zero angle of attack and for a range of subsonic and transonic Mach numbers.

Bennett, R. M.↗

Calculation of transonic steady and oscillatory pressures on a low aspect ratio model and comparison with experiment

Pressure data measured by the British Royal Aircraft Establishment for the AGARD SMP tailplane are compared with results calculated using the transonic small perturbation code XTRAN3S. A brief description of the analysis is given and a recently developed finite difference grid is described. Results are presented for five steady and nine harmonically oscillating cases near zero angle of attack and for a range of subsonic and transonic Mach numbers.

Bennett, R. M.↗

AGARD standard aeroelastic configurations for dynamic response. 1: Wing 445.6

This report contains experimental flutter data for the AGARD 3D swept tapered standard configuration "Wing 445.6", along with related descriptive data of the model properties required for comparative flutter calculations. As part of a cooperative AGARD-SMP programme, guided by the Sub-Committee on Aeroelasticity, this standard configuration may serve as a common basis for comparisons of calculated and measured aeroelastic behaviour. These comparisons will promote a better understanding of the assumptions, approximations and limitations underlying the various aerodynamic methods applied, thus pointing the way to further improvements.

Aircraft↗

Building Mathematical Models Of Solid Objects

Solid Modeling Program (SMP) version 2.0 provides capability to model complex solid objects mathematically through aggregation of geometric primitives (parts). System provides designer with basic set of primitive parts and capability to define new primitives. Six primitives included in present version: boxes, cones, spheres, paraboloids, tori, and trusses. Written in VAX/VMS FORTRAN 77.

Randall, Donald P.↗

A dynamic systems engineering methodology research study. Phase 2: Evaluating methodologies, tools, and techniques for applicability to NASA's systems projects

A study of NASA's Systems Management Policy (SMP) concluded that the primary methodology being used by the Mission Operations and Data Systems Directorate and its subordinate, the Networks Division, is very effective. Still some unmet needs were identified. This study involved evaluating methodologies, tools, and techniques with the potential for resolving the previously identified deficiencies. Six preselected methodologies being used by other organizations with similar development problems were studied. The study revealed a wide range of significant differences in structure. Each system had some strengths but none will satisfy all of the needs of the Networks Division. Areas for improvement of the methodology being used by the Networks Division are listed with recommendations for specific action.

Paul, Arthur S.↗

The changing spectrum of the LMC planetary N66

Recent spectroscopy and photometry of the planetary nebula N66 (SMP 83) in the Large Magellanic Cloud show a continuing evolution, with a central WR spectrum becoming more visible. The planetary nebula shell and Wolf-Rayet (WR) star have velocities which differ by approximately 240 km/s. Properties of this interesting object are reviewed, and we discuss its possible X-ray detection by ROSAT.

Cowley, A. P.↗

Petabyte Class Storage at Jefferson Lab (CEBAF)

By 1997, the Thomas Jefferson National Accelerator Facility will collect over one Terabyte of raw information per day of Accelerator operation from three concurrently operating Experimental Halls. When post-processing is included, roughly 250 TB of raw and formatted experimental data will be generated each year. By the year 2000, a total of one Petabyte will be stored on-line. Critical to the experimental program at Jefferson Lab (JLab) is the networking and computational capability to collect, store, retrieve, and reconstruct data on this scale. The design criteria include support of a raw data stream of 10-12 MB/second from Experimental Hall B, which will operate the CEBAF (Continuous Electron Beam Accelerator Facility) Large Acceptance Spectrometer (CLAS). Keeping up with this data stream implies design strategies that provide storage guarantees during accelerator operation, minimize the number of times data is buffered allow seamless access to specific data sets for the researcher, synchronize data retrievals with the scheduling of postprocessing calculations on the data reconstruction CPU farms, as well as support the site capability to perform data reconstruction and reduction at the same overall rate at which new data is being collected. The current implementation employs state-of-the-art StorageTek Redwood tape drives and robotics library integrated with the Open Storage Manager (OSM) Hierarchical Storage Management software (Computer Associates, International), the use of Fibre Channel RAID disks dual-ported between Sun Microsystems SMP servers, and a network-based interface to a 10,000 SPECint92 data processing CPU farm. Issues of efficiency, scalability, and manageability will become critical to meet the year 2000 requirements for a Petabyte of near-line storage interfaced to over 30,000 SPECint92 of data processing power.

Chambers, Rita↗

Implementation of Helioseismic Data Reduction and Diagnostic Techniques on Massively Parallel Architectures

Under the direction of Dr. Rhodes, and the technical supervision of Dr. Korzennik, the data assimilation of high spatial resolution solar dopplergrams has been carried out throughout the program on the Intel Delta Touchstone supercomputer. With the help of a research assistant, partially supported by this grant, and under the supervision of Dr. Korzennik, code development was carried out at SAO, using various available resources. To ensure cross-platform portability, PVM was selected as the message passing library. A parallel implementation of power spectra computation for helioseismology data reduction, using PVM was successfully completed. It was successfully ported to SMP architectures (i.e. SUN), and to some MPP architectures (i.e. the CM5). Due to limitation of the implementation of PVM on the Cray T3D, the port to that architecture was not completed at the time.

Korzennik, Sylvain↗

Second International Workshop on Software Engineering and Code Design in Parallel Meteorological and Oceanographic Applications

This report contains the abstracts and technical papers from the Second International Workshop on Software Engineering and Code Design in Parallel Meteorological and Oceanographic Applications, held June 15-18, 1998, in Scottsdale, Arizona. The purpose of the workshop is to bring together software developers in meteorology and oceanography to discuss software engineering and code design issues for parallel architectures, including Massively Parallel Processors (MPP's), Parallel Vector Processors (PVP's), Symmetric Multi-Processors (SMP's), Distributed Shared Memory (DSM) multi-processors, and clusters. Issues to be discussed include: (1) code architectures for current parallel models, including basic data structures, storage allocation, variable naming conventions, coding rules and styles, i/o and pre/post-processing of data; (2) designing modular code; (3) load balancing and domain decomposition; (4) techniques that exploit parallelism efficiently yet hide the machine-related details from the programmer; (5) tools for making the programmer more productive; and (6) the proliferation of programming models (F--, OpenMP, MPI, and HPF).

OKeefe, Matthew↗

A New Compendium of Unsteady Aerodynamic Test Cases for CFD: Summary of AVT WG-003 Activities

With the continuous progress in hardware and numerical schemes, Computational Unsteady Aerodynamics (CUA), that is, the application of Computational Fluid Dynamics (CFD) to unsteady flowfields, is slowly finding its way as a useful and reliable tool (turbulence and transition modeling permitting) in the aircraft, helicopter, engine and missile design and development process. Before a specific code may be used with confidence it is essential to validate its capability to describe the physics of the flow correctly, or at least to the level of approximation required, for which purpose a comparison with accurate experimental data is needed. Unsteady wind tunnel testing is difficult and expensive; two factors which dramatically limit the number of organizations with the capability and/or resources to perform it. Thus, unsteady experimental data is scarce, often classified and scattered in diverse documents. Additionally, access to the reports does not necessarily assure access to the data itself. The collaborative effort described in this paper was conceived with the aim of collecting into a single easily accessible document as much quality data as possible. The idea is not new. In the early 80's NATO's AGARD (Advisory Group for Aerospace Research & Development) Structures and Material Panel (SMP) produced AGARD Report No. 702 "Compendium of Unsteady Aerodynamic Measurements", which has found and continues to find extensive use within the CUA Community. In 1995 AGARD's Fluid Dynamics Panel (FDP) decided to update and expand the former database with new geometries and physical phenomena, and launched Working Group WG-22 on "Validation Data for Computational Unsteady Aerodynamic Codes". Shortly afterwards AGARD was reorganized as the RTO (Research and Technology Organization) and the WG was renamed as AVT (Applied Vehicle Technolology) WG-003. Contributions were received from AEDC, BAe, DLR, DERA, Glasgow University, IAR, NAL, NASA, NLR, and ONERA. The final publication with the results of the exercise is expected in the second part of 1999. The aim of the present paper is to announce and present the new database to the Aeroelasticity community. It is also intended to identify, together with one of the groups of end users it targets, deficiencies in the compendium that should be addressed by means of new wind tunnel tests or by obtaining access to additionally existing data.

Ruiz-Calavera, Luis P.↗

Message Passing vs. Shared Address Space on a Cluster of SMPs

The convergence of scalable computer architectures using clusters of PCs (or PC-SMPs) with commodity networking has become an attractive platform for high end scientific computing. Currently, message-passing and shared address space (SAS) are the two leading programming paradigms for these systems. Message-passing has been standardized with MPI, and is the most common and mature programming approach. However message-passing code development can be extremely difficult, especially for irregular structured computations. SAS offers substantial ease of programming, but may suffer from performance limitations due to poor spatial locality, and high protocol overhead. In this paper, we compare the performance of and programming effort, required for six applications under both programming models on a 32 CPU PC-SMP cluster. Our application suite consists of codes that typically do not exhibit high efficiency under shared memory programming. due to their high communication to computation ratios and complex communication patterns. Results indicate that SAS can achieve about half the parallel efficiency of MPI for most of our applications: however, on certain classes of problems SAS performance is competitive with MPI. We also present new algorithms for improving the PC cluster performance of MPI collective operations.

Shan, Hongzhang↗

Performance Characteristics of the Multi-Zone NAS Parallel Benchmarks

We describe a new suite of computational benchmarks that models applications featuring multiple levels of parallelism. Such parallelism is often available in realistic flow computations on systems of grids, but had not previously been captured in bench-marks. The new suite, named NPB Multi-Zone, is extended from the NAS Parallel Benchmarks suite, and involves solving the application benchmarks LU, BT and SP on collections of loosely coupled discretization meshes. The solutions on the meshes are updated independently, but after each time step they exchange boundary value information. This strategy provides relatively easily exploitable coarse-grain parallelism between meshes. Three reference implementations are available: one serial, one hybrid using the Message Passing Interface (MPI) and OpenMP, and another hybrid using a shared memory multi-level programming model (SMP+OpenMP). We examine the effectiveness of hybrid parallelization paradigms in these implementations on three different parallel computers. We also use an empirical formula to investigate the performance characteristics of the multi-zone benchmarks.

Jin, Haoqiang↗

Alloy Design Workbench-Surface Modeling Package Developed

NASA Glenn Research Center's Computational Materials Group has integrated a graphical user interface with in-house-developed surface modeling capabilities, with the goal of using computationally efficient atomistic simulations to aid the development of advanced aerospace materials, through the modeling of alloy surfaces, surface alloys, and segregation. The software is also ideal for modeling nanomaterials, since surface and interfacial effects can dominate material behavior and properties at this level. Through the combination of an accurate atomistic surface modeling methodology and an efficient computational engine, it is now possible to directly model these types of surface phenomenon and metallic nanostructures without a supercomputer. Fulfilling a High Operating Temperature Propulsion Components (HOTPC) project level-I milestone, a graphical user interface was created for a suite of quantum approximate atomistic materials modeling Fortran programs developed at Glenn. The resulting "Alloy Design Workbench-Surface Modeling Package" (ADW-SMP) is the combination of proven quantum approximate Bozzolo-Ferrante-Smith (BFS) algorithms (refs. 1 and 2) with a productivity-enhancing graphical front end. Written in the portable, platform independent Java programming language, the graphical user interface calls on extensively tested Fortran programs running in the background for the detailed computational tasks. Designed to run on desktop computers, the package has been deployed on PC, Mac, and SGI computer systems. The graphical user interface integrates two modes of computational materials exploration. One mode uses Monte Carlo simulations to determine lowest energy equilibrium configurations. The second approach is an interactive "what if" comparison of atomic configuration energies, designed to provide real-time insight into the underlying drivers of alloying processes.

Abel, Phillip B.↗