Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “high throughput computing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Concepts leading to the IMAGE-100 hybrid interactive system

As LACIE Procedure 1 evolved from the Classification and Mensuration Subsystem smallfields procedures, it became evident that two computational systems would have merit-the LACIE/Earth Resources Interactive Processing System based on a large IBM-360 computer oriented for operational use with high computational throughput, and a smaller, highly interactive system based on a PDP 11-45 minicomputer and its display system, the IMAGE-100. The latter had advantages for certain phases; notably, interactive spectral aids could be implemented quite rapidly. This would allow testing and development of Procedure 1 before its implementation on the LACIE/Earth Resources Interactive Processing System. The resulting minicomputer system, called the Classification and Mensuration Subsystem IMAGE-100 Hybrid System, allowed Procedure-1 operations to be performed interactively, except for clustering, classification, and automatic selection of best acquisitions, which were offloaded to the LACIE/Earth Resources Interactive Processing System.

Mackin, T. F.↗

Advanced information processing system: The Army fault tolerant architecture conceptual study. Volume 1: Army fault tolerant architecture overview

Digital computing systems needed for Army programs such as the Computer-Aided Low Altitude Helicopter Flight Program and the Armored Systems Modernization (ASM) vehicles may be characterized by high computational throughput and input/output bandwidth, hard real-time response, high reliability and availability, and maintainability, testability, and producibility requirements. In addition, such a system should be affordable to produce, procure, maintain, and upgrade. To address these needs, the Army Fault Tolerant Architecture (AFTA) is being designed and constructed under a three-year program comprised of a conceptual study, detailed design and fabrication, and demonstration and validation phases. Described here are the results of the conceptual study phase of the AFTA development. Given here is an introduction to the AFTA program, its objectives, and key elements of its technical approach. A format is designed for representing mission requirements in a manner suitable for first order AFTA sizing and analysis, followed by a discussion of the current state of mission requirements acquisition for the targeted Army missions. An overview is given of AFTA's architectural theory of operation.

Harper, R. E.↗

Computational Approaches for Li-O2 Battery Design

Threshold energy densities for general aviation electric aircraft are 400 Wh/kg with more ambitious air vehicles having significantly higher requirements. Li-O2 batteries, with the highest theoretical capacity, are one of the few “beyond Li-ion” chemistries that might satisfy the extraordinary specific capacity as well as specific power requirements of electric aircraft. However, side reactions at interfaces, in particular at the cathode, over the charge-discharge cycles result in very short cycle-life and dramatic reduction of capacity. Addressing these issues, in addition to others, are crucial for realizing practical, high performance Li-O2 batteries. In this talk, we discuss atomistic computational work to understand and mitigate some of the issues, including those at interfaces, that affect the Li-O2 electrochemistry. To start, we discuss the deposition mechanisms, both surface and solution based, of Li2O2 and their dependence on external potential. Next, we explore molten salt electrolytes as a stable alternative to organic electrolytes and approaches taken to develop new practical molten salt eutectic mixtures. Finally, we address the issue of reactive carbon-cathodes and possible cathode-candidates that were identified throughput high-throughput computations.

Li-Air battery↗

Using Ada for a distributed, fault tolerant system

It is pointed out that advanced avionics applications increasingly require underlying machine architectures which are damage and fault tolerant, and which provide access to distributed sensors, effectors and high-throughput computational resources. The Advanced Information Processing System (AIPS), sponsored by NASA, is to provide an architecture which can meet the considered requirements. Ada was selected for implementing the AIPS system software. Advantages of Ada are related to its provisions for real-time programming, error detection, modularity and separate compilation, and standardization and portability. Chief drawbacks of this language are currently limited availability and maturity of language implementations, and limited experience in applying the language to real-time applications. The present investigation is concerned with current plans for employing Ada in the design of the software for AIPS. Attention is given to an overview of AIPS, AIPS software services, and representative design issues in each of four major software categories.

Dewolf, J. B.↗

Designing Molten Salt Eutectics: A Combined Thermodynamic Modeling and Machine Learning Approach

Designing stable electrolytes with target properties is an important challenge in realizing next generation energy storage devices. Molten salt eutectics-based electrolytes are known for their stability with minimal parasitic reactions when compared to traditional organic electrolytes and are an attractive option for different battery chemistries. The operating temperature of the molten salt batteries depends on the melting temperature of the eutectic and hence there is a necessity to discover novel low melting temperature molten salt eutectic mixtures for energy storage applications. In this work we develop a high throughput computational screening approach for molten salt mixtures using thermodynamic modeling and machine learning (ML). COSMO-SAC model and ML approaches were independently developed based on the existing experimental data and these models were further used to predict the eutectic melting temperature and composition of several new binary, ternary, and quaternary mixtures. We show that combining ML and thermodynamic modeling strategies is effective in exploring the vast design space of molten salt mixtures.

Thermodynamics↗

Designing Molten Salt Eutectics

Designing stable electrolytes with target properties is an important challenge in realizing next generation energy storage devices. Molten salt eutectics-based electrolytes are known for their stability with minimal parasitic reactions when compared to traditional organic electrolytes and are an attractive option for different battery chemistries. The operating temperature of the molten salt batteries depends on the melting temperature of the eutectic and hence there is a necessity to discover novel low melting temperature molten salt eutectic mixtures for energy storage applications. In this work we develop a high throughput computational screening approach for molten salt mixtures using thermodynamic modeling and machine learning (ML). COSMO-SAC model and ML approaches were independently developed based on the existing experimental data and these models were further used to predict the eutectic melting temperature and composition of several new binary, ternary, and quaternary mixtures. We show that combining ML and thermodynamic modeling strategies is effective in exploring the vast design space of molten salt mixtures.

Ashwin Ravichandran↗

ICME for NASA Aerospace Applications: Batteries for Electric Aviation

NASA’s approach to computational materials modeling is detailed in the NASA Vision 2040 Roadmap for Multiscale Modeling and Simulation of Materials and Systems. This report is in the spirit of national initiatives such as the Material Genome Initiative (MGI), Integrated Computational Materials Engineering (ICME), and others. We utilize a combination of fundamental modeling, computational high-throughput screening, and data science methods, e.g., machine learning, are used to find innovative solutions to NASA or national technology challenges. Applications of interest are wide ranging from advanced alloys to batteries to coatings, among others. In this talk, we present three examples for recent work related to NASA applications. First, doping advanced sulfur battery cathodes with selenium boosts electrical conductivity important for electric aircraft applications. First principles calculations will be discussed that result in compositional design maps for these materials. Second, development of icephobic coatings is important to mitigate safety hazards associated with icing for aircraft. Molecular dynamics simulations are reported for ice-surface interfaces to understand adhesion mechanisms and help screen optimal ice-phobic coatings. Third, shape memory alloys have numerous applications as actuators, superelastic materials, etc. for aerospace. We report machine learning models that predict martensitic transition temperatures across a broad swath of compositional space.

John Lawson↗

A Low-Power High-Speed Smart Sensor Design for Space Exploration Missions

A low-power high-speed smart sensor system based on a large format active pixel sensor (APS) integrated with a programmable neural processor for space exploration missions is presented. The concept of building an advanced smart sensing system is demonstrated by a system-level microchip design that is composed with an APS sensor, a programmable neural processor, and an embedded microprocessor in a SOI CMOS technology. This ultra-fast smart sensor system-on-a-chip design mimics what is inherent in biological vision systems. Moreover, it is programmable and capable of performing ultra-fast machine vision processing in all levels such as image acquisition, image fusion, image analysis, scene interpretation, and control functions. The system provides about one tera-operation-per-second computing power which is a two order-of-magnitude increase over that of state-of-the-art microcomputers. Its high performance is due to massively parallel computing structures, high data throughput rates, fast learning capabilities, and advanced VLSI system-on-a-chip implementation.

Fang, Wai-Chi↗

NASA Tech Briefs, August 2012

Topics covered include: Mars Science Laboratory Drill; Ultra-Compact Motor Controller; A Reversible Thermally Driven Pump for Use in a Sub-Kelvin Magnetic Refrigerator; Shape Memory Composite Hybrid Hinge; Binding Causes of Printed Wiring Assemblies with Card-Loks; Coring Sample Acquisition Tool; Joining and Assembly of Bulk Metallic Glass Composites Through Capacitive Discharge; 670-GHz Schottky Diode-Based Subharmonic Mixer with CPW Circuits and 70-GHz IF; Self-Nulling Lock-in Detection Electronics for Capacitance Probe Electrometer; Discontinuous Mode Power Supply; Optimal Dynamic Sub-Threshold Technique for Extreme Low Power Consumption for VLSI; Hardware for Accelerating N-Modular Redundant Systems for High-Reliability Computing; Blocking Filters with Enhanced Throughput for X-Ray Microcalorimetry; High-Thermal-Conductivity Fabrics; Imidazolium-Based Polymeric Materials as Alkaline Anion-Exchange Fuel Cell Membranes; Electrospun Nanofiber Coating of Fiber Materials: A Composite Toughening Approach; Experimental Modeling of Sterilization Effects for Atmospheric Entry Heating on Microorganisms; Saliva Preservative for Diagnostic Purposes; Hands-Free Transcranial Color Doppler Probe; Aerosol and Surface Parameter Retrievals for a Multi-Angle, Multiband Spectrometer LogScope; TraceContract; AIRS Maps from Space Processing Software; POSTMAN: Point of Sail Tacking for Maritime Autonomous Navigation; Space Operations Learning Center; OVERSMART Reporting Tool for Flow Computations Over Large Grid Systems; Large Eddy Simulation (LES) of Particle-Laden Temporal Mixing Layers; Projection of Stabilized Aerial Imagery Onto Digital Elevation Maps for Geo-Rectified and Jitter-Free Viewing; Iterative Transform Phase Diversity: An Image-Based Object and Wavefront Recovery; 3D Drop Size Distribution Extrapolation Algorithm Using a Single Disdrometer; Social Networking Adapted for Distributed Scientific Collaboration; General Methodology for Designing Spacecraft Trajectories; Hemispherical Field-of-View Above-Water Surface Imager for Submarines; and Quantum-Well Infrared Photodetector (QWIP) Focal Plane Assembly.

Source record↗

Strain-Layer-Superlattice Light Modulator

Conceptual device combines resonant reflection and photovoltaic action to enable one light beam to impose spatial and temporal modulation on another light beam. Such spatial light modulator, with high speed and multiplicity of parallel signal channels, used in image processing or similar computation requiring high data-throughput rates. Microstructures of GaAs and InAs with multiple quantum wells and compositional superlattices grown by molecular-beam epitaxy. Enhanced electro-optical properties of arrangement of alternating layers enables writing light beam to modulate reading light beam.

Maserjian, Joseph↗

Parallelism and pipelining in high-speed digital simulators

The attainment of high computing speed as measured by the computational throughput is seen as one of the most challenging requirements. It is noted that high speed is cardinal in several distinct classes of applications. These classes are then discussed; they comprise (1) the real-time simulation of dynamic systems , (2) distributed parameter systems, and (3) mixed lumped and distributed systems. From the 1950s on, the quest for high speed in digital simulators concentrated on overcoming the limitations imposed by the so-called von Neumann bottleneck. Two major architectural approaches have made ig possible to circumvent this bottleneck and attain high speeds. These are pipelining and parallelism. Supercomputers, peripheral array processors, and microcomputer networks are then discussed.

Karplus, W. J.↗

Informatics and High Throughput Screening of Thermophysical Properties

The combination of computer-aided experiments with computational modeling enables a new class of powerful tools for materials research. A non-contact method for measuring density, thermal expansion, and creep of undercooled and high-temperature materials has been developed, using electrostatic levitation and optical diagnostics, including digital video. These experiments were designed to take advantage of the large volume of data (many gigabytes/experiment, terabytes/campaign) to gain additional information about the samples. For example, using sub-pixel interpolation to measure about 1000 vectors per image of the sample's surface allows the density of an axisymmetric sample to be determined to an accuracy of about 200 ppm (0.02%). A similar analysis applied to the surface shape of a rapidly rotating sample is combined with finite element modeling to determine the stress-dependence of creep in the sample in a single test. Details of the methods for both the computer-aided experiments and computational models will be discussed.

Hyers, Robert W.↗

Implicit Thermochemical Nonequilibrium Flow Simulations on Unstructured Grids using GPUs

Thermochemical nonequilibrium flow simulation capabilities have been previously implemented, verified, and validated for central processing unit (CPU) systems in NASA’s unstructured-grid computational fluid dynamics solver FUN3D. Many exascale-class high-performance computing systems will rely on graphics processing unit (GPU) architectures for high throughput and energy efficiency; thus, CPU-based scientific computing software unable to effectively utilize these systems must be updated. In this work, we present a CUDA C++ implementation of FUN3D’s thermochemical nonequilibrium flow simulation capabilities targeting NVIDIA Tesla GPUs. An overview of the porting and optimization strategy is described and performance comparisons with other recent architectures are presented. Scaling to thousands of GPUs is demonstrated, yielding computational performance equivalent to that of several million CPU cores. The implementation enables efficient, high-fidelity, scale-resolving simulations of thermochemical nonequilibrium flows for many applications including atmospheric entry, hypersonics, and combustion.

GPU↗

Analog Correlator Based on One Bit Digital Correlator

A two input time domain correlator may perform analog correlation. In order to achieve high throughput rates with reduced or minimal computational overhead, the input data streams may be hard limited through adaptive thresholding to yield two binary bit streams. Correlation may be achieved through the use of a Hamming distance calculation, where the distance between the two bit streams approximates the time delay that separates them. The resulting Hamming distance approximates the correlation time delay with high accuracy.

Prokop, Norman↗

High performance remote sensing data analysis using parallel computation

This paper examines the JPL/Caltech parallel processing system designed for rapid processing and transfer of large quantities of data from remote sensing instruments flown on NASA missions. Two remote sensing analysis applications that use this processing system are described: (1) an analysis system for retrieval of atmospheric parameters (such as species abundance, atmospheric temperature, and water vapor profiles) from data obtained by a Fourier transform IR spectrometer and (2) a prototype airborne SAR processing system. It is shown that a parallel processing system such as the JPL/Caltech system can offer supercomputer computational capability and high-volume data throughput and still be cost-effective.

Patterson, Jean E.↗

Modeling and Simulation Reliable Spacecraft On-Board Computing

The proposed project will investigate modeling and simulation-driven testing and fault tolerance schemes for Spacecraft On-Board Computing, thereby achieving reliable spacecraft telecommunication. A spacecraft communication system has inherent capabilities of providing multipoint and broadcast transmission, connectivity between any two distant nodes within a wide-area coverage, quick network configuration /reconfiguration, rapid allocation of space segment capacity, and distance-insensitive cost. To realize the capabilities above mentioned, both the size and cost of the ground-station terminals have to be reduced by using reliable, high-throughput, fast and cost-effective on-board computing system which has been known to be a critical contributor to the overall performance of space mission deployment. Controlled vulnerability of mission data (measured in sensitivity), improved performance (measured in throughput and delay) and fault tolerance (measured in reliability) are some of the most important features of these systems. The system should be thoroughly tested and diagnosed before employing a fault tolerance into the system. Testing and fault tolerance strategies should be driven by accurate performance models (i.e. throughput, delay, reliability and sensitivity) to find an optimal solution in terms of reliability and cost. The modeling and simulation tools will be integrated with a system architecture module, a testing module and a module for fault tolerance all of which interacting through a centered graphical user interface.

Park, Nohpill↗

A modular minicomputer based Navier-Stokes solver

The basic module consists of a minicomputer, low cost peripheral storage device (disk) and a modest number (8-12) of microcomputer modules. A simple arrangement, where the microcomputers are connected to a single time multiplexed bus, only communicating to the host minicomputer, will be efficient. By running the machine in a dedicated mode for long periods of time, it will be possible to obtain a large number of solutions. As such, the device should be useful as a research tool. A scheme is outlined to assemble a number of these computing modules in parallel to decrease computing time. The advantages and disadvantages are discussed of using a number of these systems assembled in a loosely coupled configuration, each independently computing a separate flow, to give a very high throughput.

Steinhoff, J.↗

Exploring the use of I/O nodes for computation in a MIMD multiprocessor

As parallel systems move into the production scientific-computing world, the emphasis will be on cost-effective solutions that provide high throughput for a mix of applications. Cost effective solutions demand that a system make effective use of all of its resources. Many MIMD multiprocessors today, however, distinguish between 'compute' and 'I/O' nodes, the latter having attached disks and being dedicated to running the file-system server. This static division of responsibilities simplifies system management but does not necessarily lead to the best performance in workloads that need a different balance of computation and I/O. Of course, computational processes sharing a node with a file-system service may receive less CPU time, network bandwidth, and memory bandwidth than they would on a computation-only node. In this paper we begin to examine this issue experimentally. We found that high performance I/O does not necessarily require substantial CPU time, leaving plenty of time for application computation. There were some complex file-system requests, however, which left little CPU time available to the application. (The impact on network and memory bandwidth still needs to be determined.) For applications (or users) that cannot tolerate an occasional interruption, we recommend that they continue to use only compute nodes. For tolerant applications needing more cycles than those provided by the compute nodes, we recommend that they take full advantage of both compute and I/O nodes for computation, and that operating systems should make this possible.

Kotz, David↗