Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “shared memory”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 343 records · Page 19

Job Management Requirements for NAS Parallel Systems and Clusters

A job management system is a critical component of a production supercomputing environment, permitting oversubscribed resources to be shared fairly and efficiently. Job management systems that were originally designed for traditional vector supercomputers are not appropriate for the distributed-memory parallel supercomputers that are becoming increasingly important in the high performance computing industry. Newer job management systems offer new functionality but do not solve fundamental problems. We address some of the main issues in resource allocation and job scheduling we have encountered on two parallel computers - a 160-node IBM SP2 and a cluster of 20 high performance workstations located at the Numerical Aerodynamic Simulation facility. We describe the requirements for resource allocation and job management that are necessary to provide a production supercomputing environment on these machines, prioritizing according to difficulty and importance, and advocating a return to fundamental issues.

Saphir, William↗

Crystallographic and general use programs for the XDS Sigma 5 computer

Programs in basic FORTRAN 4 are described, which fall into three catagories: (1) interactive programs to be executed under time sharing (BTM); (2) non interactive programs which are executed in batch processing mode (BPM); and (3) large non interactive programs which require more memory than is available in the normal BPM/BTM operating system and must be run overnight on a special system called XRAY which releases about 45,000 words of memory to the user. Programs in catagories (1) and (2) are stored as FORTRAN source files in the account FSNYDER. Programs in catagory (3) are stored in the XRAY system as load modules. The type of file in account FSNYDER is identified by the first two letters in the name.

Snyder, R. L.↗

Preface to the Special Collection: Recollections in Space Physics

Space Physics is a comparatively young scientific discipline, tightly linked to the era of satellite based investigations and the discoveries that came with it. As such, we as a community are fortunate to have met, been taught and mentored by, and even become friends with, many of those who were around to witness the birth of, and in many cases establish, our field. Each of these "Pioneers" have remarkable stories to tell. These accounts from the dawn of the space age provide a glimpse into that era of scientific discovery and are an important part of our collective history that deserve to be shared and commemorated. These are stories of perseverance, ingenuity, luck, and sometimes failure. We are fortunate to live in an era in which these distinguished scientists are still with us and able to share their experiences.To commemorate the 100th Anniversary of the American Geophysical Union (AGU), we have solicited a special collection of recollection papers to memorialize these experiences. This is not the first such effort to capture these stories. In 1994 (Vol. 99, No. A10, pp. 19,099-19,212) and again in 1996 (Vol. 101, No. A5, pp. 10,477-10,585) JGRSpace Physics published special sections entitled, "Pioneers of Space Physics," in celebration of AGU's 75th anniversary. Later, in 1997, Gilmore and Spreiter (1997) added a set of recollection papers on the discovery of the magnetosphere. Together, these papers containing personal accounts from the pioneers of the space age covered the period of roughly 1958-1967.It is our great honor to present retrospective papers from several of our distinguished colleagues who contributed significantly to the explosion in understanding of space physics and aeronomy that occurred from roughly 1967-1980. We chose to start at the end of the previous set of recollections and cover the decade of the 1970s, during which our field greatly expanded. It is also a time in which the second generation of space scientists entered the field and made foundational discoveries that continue to define our discipline.As with the previous papers, we asked the authors to "recount some of the events leading to the emergence of space physics and to put the events into a professional as well as personal perspective." (Gombosi et al.,1994). The authors of this special collection of recollections were selected following the same guidelines as the previous efforts. We started from a long list of senior colleagues and narrowed the list to roughly two dozen distinguished scientists based on discipline and geographic balance. Some of our colleagues, when asked, felt they would be unable to devote the time or energy to such a recollection and declined. Sadly,in the intervening 25 years since the original effort, some of our colleagues who were most active during the early years of the space age have passed. Therefore, this new special collection, 25 years after the first one, is timely with the AGU centennial celebrations, but late in fully capturing the stories of the pioneers of space physics. This is unfortunate and we encourage future editors of this journal to commission specialsections of legacy perspectives with a faster cadence than a quarter of a century.

Kepko, Emil L.↗

Improvements to a Batch Pentadiagonal Solver on NVIDIA GPUs

This poster presents the recent work in OVERFLOW to port the batched pentadiagonal solver to NVIDIA GPUs. There are five pentadiagonal systems for each pencil in the grid but three of these systems share the same LHS. Our first simple approach for porting the pentadiagonal solver to the GPUs was to take advantage of the shared LHS by assigning three threads to the three LHS of each pencil. We demonstrated that this custom solver was 92% faster than the NVIDIA batched pentadiagonal library implementation on a V100 GPU due to the lower memory bandwidth requirements. The second approach treated each pentadiagonal system as a 2x2 block tridiagonal system and used a variant of the parallel cyclic reduction algorithm to solve the problem. One benefit of this approach is that it does not require interleaving the data between each system. We demonstrated that this algorithm is 2.18x faster than the NVIDIA library implementation for the same amount of work. If we take advantage of our shared LHS, this approach is 2.58x faster than the library implementation on a V100 GPU.

GPU Programming↗

Jell-Molds and Cookie-cutters: Shrinkwrap Isn't Just for Leftovers Anymore

So what is Shrinkwrap all about? For those of you who may not know about it, Shrinkwrap is a type of data structure that can manifest itself as a feature or model. It is cleverly covered up, almost hidden, and doesn't get the press or widespread use of a solid or surface. The shrinkwrap feature is located under the data sharing submenu of the feature menu. The shrinkwrap feature, as described by PTC, is a collection of surfaces and datum features of a model that represents the exterior of the model . The advantages and applications of the shrinkwrap feature are in the creation of minimal memory guzzling representations of assemblies. These can be used to represent subassemblies in parent assemblies, and can handle control of dependency issues, geometry represented, and additional references through the use of the shrinkwrap feature options. The shrinkwrap model is an option available under the save as umbrella. Its function, as described by PTC, is to share data with internal and external design groups and improve performance in large assembly design . Some of the benefits of the shrinkwrap model include being able to represent complex assemblies with a single, lightweight part that protects design intent and parametric data, and the ability to improve performance of large assembly modeling in the area of less load time. The proper-scale models can be saved as IGES, STEP, and VRML (for fly-throughs).

Randazzo, John↗

The effects of stimulus modality and task integrality: Predicting dual-task performance and workload from single-task levels

The influence of stimulus modality and task difficulty on workload and performance was investigated. The goal was to quantify the cost (in terms of response time and experienced workload) incurred when essentially serial task components shared common elements (e.g., the response to one initiated the other) which could be accomplished in parallel. The experimental tasks were based on the Fittsberg paradigm; the solution to a SternBERG-type memory task determines which of two identical FITTS targets are acquired. Previous research suggested that such functionally integrated dual tasks are performed with substantially less workload and faster response times than would be predicted by suming single-task components when both are presented in the same stimulus modality (visual). The physical integration of task elements was varied (although their functional relationship remained the same) to determine whether dual-task facilitation would persist if task components were presented in different sensory modalities. Again, it was found that the cost of performing the two-stage task was considerably less than the sum of component single-task levels when both were presented visually. Less facilitation was found when task elements were presented in different sensory modalities. These results suggest the importance of distinguishing between concurrent tasks that complete for limited resources from those that beneficially share common resources when selecting the stimulus modalities for information displays.

Hart, S. G.↗

The effects of syntactic complexity on the human-computer interaction

Three divided-attention experiments were performed to evaluate the effectiveness of a syntactic analysis of the primary task of editing flight route-way-point information. For all editing conditions, a formal syntactic expression was developed for the operator's interaction with the computer. In terms of the syntactic expression, four measures of syntactic were examined. Increased syntactic complexity did increase the time to train operators, but once the operators were trained, syntactic complexity did not influence the divided-attention performance. However, the number of memory retrievals required of the operator significantly accounted for the variation in the accuracy, workload, and task completion time found on the different editing tasks under attention-sharing conditions.

Chechile, R. A.↗

General-purpose interface bus for multiuser, multitasking computer system

The architecture of a multiuser, multitasking, virtual-memory computer system intended for the use by a medium-size research group is described. There are three central processing units (CPU) in the configuration, each with 16 MB memory, and two 474 MB hard disks attached. CPU 1 is designed for data analysis and contains an array processor for fast-Fourier transformations. In addition, CPU 1 shares display images viewed with the image processor. CPU 2 is designed for image analysis and display. CPU 3 is designed for data acquisition and contains 8 GPIB channels and an analog-to-digital conversion input/output interface with 16 channels. Up to 9 users can access the third CPU simultaneously for data acquisition. Focus is placed on the optimization of hardware interfaces and software, facilitating instrument control, data acquisition, and processing.

Generazio, Edward R.↗

Autonomous Information Unit for Fine-Grain Data Access Control and Information Protection in a Net-Centric System

As communication and networking technologies advance, networks will become highly complex and heterogeneous, interconnecting different network domains. There is a need to provide user authentication and data protection in order to further facilitate critical mission operations, especially in the tactical and mission-critical net-centric networking environment. The Autonomous Information Unit (AIU) technology was designed to provide the fine-grain data access and user control in a net-centric system-testing environment to meet these objectives. The AIU is a fundamental capability designed to enable fine-grain data access and user control in the cross-domain networking environments, where an AIU is composed of the mission data, metadata, and policy. An AIU provides a mechanism to establish trust among deployed AIUs based on recombining shared secrets, authentication and verify users with a username, X.509 certificate, enclave information, and classification level. AIU achieves data protection through (1) splitting data into multiple information pieces using the Shamir's secret sharing algorithm, (2) encrypting each individual information piece using military-grade AES-256 encryption, and (3) randomizing the position of the encrypted data based on the unbiased and memory efficient in-place Fisher-Yates shuffle method. Therefore, it becomes virtually impossible for attackers to compromise data since attackers need to obtain all distributed information as well as the encryption key and the random seeds to properly arrange the data. In addition, since policy can be associated with data in the AIU, different user access and data control strategies can be included. The AIU technology can greatly enhance information assurance and security management in the bandwidth-limited and ad hoc net-centric environments. In addition, AIU technology can be applicable to general complex network domains and applications where distributed user authentication and data protection are necessary. AIU achieves fine-grain data access and user control, reducing the security risk significantly, simplifying the complexity of various security operations, and providing the high information assurance across different network domains.

Chow, Edward T.↗

Spectral transmissometer and radiometer - Design and initial results

A new solid-state spectral transmissometer and radiometer is described. The radiometer measures upwelling radiance, downwelling irradiance, and beam transmittance from 390 to 750 nm with channel widths of 2.35 nm. The spectrometer consists of a 256 element CCD linear array collecting light dispersed by a reflection grating in a modified Littrow configuration. The spectrometer is time and space-shared among the three signal types. The instrument has been deployed as a free-drifting buoy and in the profiling mode, with data stored internally on a magnetic bubble memory or sent up a conducting cable as desired. Power can be supplied either by a detachable external battery pack or through conducting cable. The instrument has been deployed in the oligotrophic North Pacific Central Gyre and in the eutrophic Straits of Juan de Fuca, and preliminary results for each region are discussed.

Carder, Kendall L.↗

Thread Migration in the Presence of Pointers

Dynamic migration of lightweight threads supports both data locality and load balancing. However, migrating threads that contain pointers referencing data in both the stack and heap remains an open problem. In this paper we describe a technique by which threads with pointers referencing both stack and non-shared heap data can be migrated such that the pointers remain valid after migration. As a result, threads containing pointers can now be migrated between processors in a homogeneous distributed memory environment.

Cronk, David↗

A Low-Memory Spectral-Correlation Analyzer for Digital QAM-SRRC Waveforms

Cyclostationary signal processing (CSP) provides the ability to estimate received waveforms' statistical features blindly. Quadrature amplitude modulated (QAM) waveforms, when filtered by the square-root-raised cosine (SRRC) pulse shape function, have cyclic features that CSP can exploit to detect waveform parameters such as symbol rate (SR) and center frequency (CF). The estimation of these SR-CF pairs enables a cognitive radio (CR) to perform spectrum sensing techniques such as spectrum sharing and interference mitigation. Here, we investigate a field-programmable gate array (FPGA) application of a blind symbol rate-center frequency estimator. First, this study provides a background on the theory behind the cyclic spectral density function (CSD), spectral correlation analyzers (SCA), and spectrum sensing. Following this is a discussion on the motivation for CubeSat spectrum sensing. An SCA implementation for low-memory devices, such as FPGA-based CubeSat, is then describes. The paper concludes by reporting the performance characteristics of the newly developed streaming-based SCA.

FPGA↗

Journal of Gravitational Physiology, Volume 12, Number 1

The following topics were covered: System Specificity in Responsiveness to Intermittent -Gx Gravitation during Simulated Microgravity in Rats; A Brief Overview of Animal Hypergravity Studies; Neurovestibular Adaptation to Short Radius Centrifugation; Effect of Artificial Gravity with Exercise Load by Using Short-Arm Centrifuge with Bicycle Ergometer as a Countermeasure Against Disused Osteoporosis; Perception of Body Vertical in Microgravity during Parabolic Flight; Virtual Environment a Behavioral and Countermeasure Tool for Assisted Gesture in Weightlessness: Experiments during Parabolic Flight; Artificial Gravity: Physiological Perspectives for Long-Term Space Exploration; Comparison of the Effects of DL-threo-Beta-Benzyloxyaspartate on the Glutamate Release from Synaptosomes before and after Exposure of Rats to Artificial Gravity; Do Perception and Postrotatory Vestibulo-Ocular Reflex Share the Same Gravity Reference?; Vestibular Adaptation to Changing Gravity Levels and the Orientation of Listing's Plane; Compound Mechanism Hypothesis on +Gz - Induced Brain Injury and Dysfunction of Learning and Memory; Environmental Challenge Impairs Prefrontal Brain Functions; Effect of 6-Days of Support Withdrawal on Characteristics of Balance Function; Hypergravity-Induced Changes of Neuronal Activities in CA1 Region of Rat Hippocampus; Audiological Findings in Antiorthostatic Position Modelling Microgravitation; Investigating Human Cognitive Performance during Spaceflight; The Relevance of the Minimization of Torque Exchange with the Environment in Weightlessness is Confirmed by Asimulation Study; Characteristics of the Eyes Pursuit Function during Readaptation to Terrestrial Gravity after Prolonged Flights Aboard the International Space Station; Comparison of Cognitive Performance Tests for Promethazine Pharmacodynamics in Human Subjects; Structural Reappraisal of Dendritic Tree of Cerebellar Purkinje Cell for Novel Functional Modeling of Elementary Sensorimotor Adaptive Processes; Orpheus 0 G or Ear in Microgravity to Establish Symptoms Concomittant of Inner and Middle Ear and Osteoporosis in Microgravity; Understanding Visual Perception in the Perspective of Gravity; Cortical Regions Associated with Orthostatic Stress in Conscious Humans; Restoration of Central Blood Volume: Application of a Simple Concept and Simple Device to Counteract Cardiovascular Instability in Syncope and Hemorrhage; WISE-2005: Integrative Cardiovascular Responses with LBNP during 60-Day Bed Rest in Women; Intracranial Pressure Increases during Weightlessness. A Parabolic Flights Study; Lower Limb & Portal Veins Echography for Predicting Risk of Thrombosis during a 90-D Bed Rest; Calf Tissue Liquid Stowage and Muscular and Deep Vein Distension in Orthostatic Tests after a 90-Day Head Down Bed Rest; Morphology of Brain Vessels in the Tail Suspended Rats Exposed to Intermittent 2 G; Alterations in Vasoreactivity of Femoral Artery Induced by Hindlimb Unweighting are Related to the Changes of Contractile Protein in Rats; and Respiratory Sinus Arrhythmia: A Marker of Decreased Parasympathetic Modulation after Short Duration.

Fuller, Charles A.↗

Common data buffer

Time-shared interface speeds data processing in distributed computer network. Two-level high-speed scanning approach routes information to buffer, portion of which is reserved for series of "first-in, first-out" memory stacks. Buffer address structure and memory are protected from noise or failed components by error correcting code. System is applicable to any computer or processing language.

Byrne, F.↗

Arranging computer architectures to create higher-performance controllers

Techniques for integrating microprocessors, array processors, and other intelligent devices in control systems are reviewed, with an emphasis on the (re)arrangement of components to form distributed or parallel processing systems. Consideration is given to the selection of the host microprocessor, increasing the power and/or memory capacity of the host, multitasking software for the host, array processors to reduce computation time, the allocation of real-time and non-real-time events to different computer subsystems, intelligent devices to share the computational burden for real-time events, and intelligent interfaces to increase communication speeds. The case of a helicopter vibration-suppression and stabilization controller is analyzed as an example, and significant improvements in computation and throughput rates are demonstrated.

Jacklin, Stephen A.↗

Master-slave system with force feedback based on dynamics of virtual model

A master-slave system can extend manipulating and sensing capabilities of a human operator to a remote environment. But the master-slave system has two serious problems: one is the mechanically large impedance of the system; the other is the mechanical complexity of the slave for complex remote tasks. These two problems reduce the efficiency of the system. If the slave has local intelligence, it can help the human operator by using its good points like fast calculation and large memory. The authors suggest that the slave is a dextrous hand with many degrees of freedom able to manipulate an object of known shape. It is further suggested that the dimensions of the remote work space be shared by the human operator and the slave. The effect of the large impedance of the system can be reduced in a virtual model, a physical model constructed in a computer with physical parameters as if it were in the real world. A method to determine the damping parameter dynamically for the virtual model is proposed. Experimental results show that this virtual model is better than the virtual model with fixed damping.

Nojima, Shuji↗

On nonlinear finite element analysis in single-, multi- and parallel-processors

Numerical solution of nonlinear equilibrium problems of structures by means of Newton-Raphson type iterations is reviewed. Each step of the iteration is shown to correspond to the solution of a linear problem, therefore the feasibility of the finite element method for nonlinear analysis is established. Organization and flow of data for various types of digital computers, such as single-processor/single-level memory, single-processor/two-level-memory, vector-processor/two-level-memory, and parallel-processors, with and without sub-structuring (i.e. partitioning) are given. The effect of the relative costs of computation, memory and data transfer on substructuring is shown. The idea of assigning comparable size substructures to parallel processors is exploited. Under Cholesky type factorization schemes, the efficiency of parallel processing is shown to decrease due to the occasional shared data, just as that due to the shared facilities.

Utku, S.↗

Computational Model of Human and System Dynamics in Free Flight: Studies in Distributed Control Technologies

This paper presents a set of studies in full mission simulation and the development of a predictive computational model of human performance in control of complex airspace operations. NASA and the FAA have initiated programs of research and development to provide flight crew, airline operations and air traffic managers with automation aids to increase capacity in en route and terminal area to support the goals of safe, flexible, predictable and efficient operations. In support of these developments, we present a computational model to aid design that includes representation of multiple cognitive agents (both human operators and intelligent aiding systems). The demands of air traffic management require representation of many intelligent agents sharing world-models, coordinating action/intention, and scheduling goals and actions in a potentially unpredictable world of operations. The operator-model structure includes attention functions, action priority, and situation assessment. The cognitive model has been expanded to include working memory operations including retrieval from long-term store, and interference. The operator's activity structures have been developed to provide for anticipation (knowledge of the intention and action of remote operators), and to respond to failures of the system and other operators in the system in situation-specific paradigms. System stability and operator actions can be predicted by using the model. The model's predictive accuracy was verified using the full-mission simulation data of commercial flight deck operations with advanced air traffic management techniques.

Corker, Kevin M.↗