Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Performance Portability”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 415 records · Page 23

Characterizing the performance of a POPS miniaturized optical particle counter when operated on a quadcopter drone

We first validate the performance of the Portable Optical Particle Spectrometer (POPS), a small light-weight and high sensitivity optical particle counter, against a reference scanning mobility particle sizer (SMPS) for a month-long deployment in an environment dominated by biomass burning aerosols. Subsequently, we examine any biases introduced by operating the POPS on a quadcopter drone, a DJI Matrice 200 V2. We report the root mean square difference (RMSD) and mean absolute difference (MAD) in particle number concentrations (PNCs) when mounted on the UAV and operating on the ground and when hovering at 10 m. When wind speeds are low (less than 2.6 m s –1 ), we find only modest differences in the RMSDs and MADs of 5 % and 3 % when operating at 10 m altitude. When wind speeds are between 2.6 and 7.7 m s –1 the RMSDs and MADs increase to 26.2 % and 19.1 %, respectively, when operating at 10m altitude. No statistical difference in PNCs was detected when operating on the UAV in either ascent or descent. We also find size distributions of aerosols in the accumulation mode (defined by diameter, d, where 0.1 ≤ d ≤ 1 µm) are relatively consistent between measurements at the surface and measurements at 10 m altitude, while differences in the coarse mode (here defined by d > 1 µm) are universally larger. Our results suggest that the impact of the UAV rotors on the POPS PNCs are small at low wind speeds, but when operating under a higher wind speed of up to 7.6 m s –1 , larger discrepancies occur. In addition, it appears that the POPS measures sub-micron aerosol particles more accurately than super-micron aerosol particles when airborne on the UAV. These measurements lay the foundations for determining the magnitude of potential errors that might be introduced into measured aerosol particle size distributions and concentrations owing to the turbulence created by the rotors on the UAV.

47 OTHER INSTRUMENTATION↗

Time dissemination: An update

A brief description of some of the improvements in the generation and dissemination of precise time and time interval at the U.S. Naval Observation are described. Details on some of the newer hardware developed for this purpose are presented. Data from tests of two of the more significant items, a small portable clock with performance approaching that of larger units and a GPS Time Transfer Unit capable of time transfer to an accuracy of less than 100 nanoseconds, are also given.

Putkovich, K.↗

Development of a miniaturized gas chromatograph-mass spectrometer with a microbore capillary column and an array detector

A review is presented of the demonstration of a miniaturized focal plane mass spectrograph using a microbore capillary column with an array detector to measure multiple mass spectra from narrow and closely eluted gas chromatograph peaks. It is shown that the system possesses high sensitivity and a high speed of spectral mass measurements. The combination of the capillary column with the miniaturized mass spectrograph is uniquely suited for the development of a field-portable, high-performance system.

Sinha, Mahadeva P.↗

Building the electronic industry's roadmaps

JTEC panelists found a strong consistency among the electronics firms they visited: all the firms had clear visions or roadmaps for their research and development activities and had committed resources to ensure that they achieve targeted results. The overarching vision driving Japan's electronics industry is that of achieving market success through developing appealing, high-quality, low-cost consumer goods - ahead of the competition. Specifics of the vision include improving performance, quality, and portability of consumer electronics products. Such visions help Japanese companies define in detail the roadmaps they will follow to develop new and improved electronic packaging technologies.

Boulton, William R.↗

On Designing Lightweight Threads for Substrate Software

Existing user-level thread packages employ a 'black box' design approach, where the implementation of the threads is hidden from the user. While this approach is often sufficient for application-level programmers, it hides critical design decisions that system-level programmers must be able to change in order to provide efficient service for high-level systems. By applying the principles of Open Implementation Analysis and Design, we construct a new user-level threads package that supports common thread abstractions and a well-defined meta-interface for altering the behavior of these abstractions. As a result, system-level programmers will have the advantages of using high-level thread abstractions without having to sacrifice performance, flexibility or portability.

Haines, Matthew↗

Load Balancing Sequences of Unstructured Adaptive Grids

Mesh adaption is a powerful tool for efficient unstructured grid computations but causes load imbalance on multiprocessor systems. To address this problem, we have developed PLUM, an automatic portable framework for performing adaptive large-scale numerical computations in a message-passing environment. This paper makes several important additions to our previous work. First, a new remapping cost model is presented and empirically validated on an SP2. Next, our load balancing strategy is applied to sequences of dynamically adapted unstructured grids. Results indicate that our framework is effective on many processors for both steady and unsteady problems with several levels of adaption. Additionally, we demonstrate that a coarse starting mesh produces high quality load balancing, at a fraction of the cost required for a fine initial mesh. Finally, we show that the data remapping overhead can be significantly reduced by applying our heuristic processor reassignment algorithm.

Biswas, Rupak↗

Processors, Pipelines, and Protocols for Advanced Modeling Networks

Predictive capabilities arise from our understanding of natural processes and our ability to construct models that accurately reproduce these processes. Although our modeling state-of-the-art is primarily limited by existing computational capabilities, other technical areas will soon present obstacles to the development and deployment of future predictive capabilities. Advancement of our modeling capabilities will require not only faster processors, but new processing algorithms, high-speed data pipelines, and a common software engineering framework that allows networking of diverse models that represent the many components of Earth's climate and weather system. Development and integration of these new capabilities will pose serious challenges to the Information Systems (IS) technology community. Designers of future IS infrastructures must deal with issues that include performance, reliability, interoperability, portability of data and software, and ultimately, the full integration of various ES model systems into a unified ES modeling network.

Coughlan, Joseph↗

Taking evolutionary circuit design from experimentation to implementation: some useful techniques and a silicon demonstration

Current techniques in evolutionary synthesis of analogue and digital circuits designed at transistor level have focused on achieving the desired functional response, without paying sufficient attention to issues needed for a practical implementation of the resulting solution. No silicon fabrication of circuits with topologies designed by evolution has been done before, leaving open questions on the feasibility of the evolutionary circuit design approach, as well as on how high-performance, robust, or portable such designs could be when implemented in hardware. It is argued that moving from evolutionary 'design-for experimentation' to 'design-for-implementation' requires, beyond inclusion in the fitness function of measures indicative of circuit evaluation factors such as power consumption and robustness to temperature variations, the addition of certain evaluation techniques that are not common in conventional design. Several such techniques that were found to be useful in evolving designs for implementation are presented; some are general, and some are particular to the problem domain of transistor-level logic design, used here as a target application. The example used here is a multifunction NAND/NOR logic gate circuit, for which evolution obtained a creative circuit topology more compact than what has been achieved by multiplexing a NAND and a NOR gate. The circuit was fabricated in a 0.5 mum CMOS technology and silicon tests showed good correspondence with the simulations.

digital circuits designs↗

Direct Methanol Fuel Cell for Portable Applications

A five cell direct methanol fuel cell stack has been developed at the Jet Propulsion Laboratory. Presently direct methanol fuel cell technology is being incorporated into a system for portable applications. Electrochemical performance and its dependence on flow rate and temperature for the five fuel cell are presented.

fuel↗

Online Photometric Calibration of Automatic Gain Thermal Infrared Cameras

Thermal infrared cameras are increasingly being used in various applications such as robot vision, industrial inspection and medical imaging, thanks to their improved resolution and portability. However, the performance of traditional computer vision techniques developed for electro-optical imagery does not directly translate to the thermal domain due to two major reasons: these algorithms require photometric assumptions to hold, and methods for photometric calibration of RGB cameras cannot be applied to thermal-infrared cameras due to difference in data acquisition and sensor phenomenology. In this paper, we take a step in this direction, and introduce a novel algorithm for online photometric calibration of thermalinfrared cameras. Our proposed method does not require any specific driver/hardware support and hence can be applied to any commercial off-the-shelf thermal IR camera. We present this in the context of visual odometry and SLAM algorithms, and demonstrate the efficacy of our proposed system through extensive experiments for both standard benchmark datasets, and real-world field tests with a thermal-infrared camera in natural outdoor environments.

Daftry, Shreyansh↗

hippynn Python Package

hippynn is a python package for defining, training, and applying neural networks to atomistic systems. In particular, it focuses on Hierarchical Interacting Particle Neural Networks (HIP-NNs), a deep learning architecture for atomistic systems. HIP-NNs take input data describing the properties of atomistic systems (often generated using ab-initio quantum mechanics) to learn fast and accurate models. A trained HIP-NN can predict potential energy surfaces, atomic charges, and more. hippynn uses PyTorch for portable high-performance code, including both CPU and GPU support. hippynn allows for extensive customization, including user-defined models and loss functions, to facilitate future research into extensions of HIP-NN as well as other atomistic deep learning models.

Lubbers, Nicolas↗

A design methodology for portable software on parallel computers

This final report for research that was supported by grant number NAG-1-995 documents our progress in addressing two difficulties in parallel programming. The first difficulty is developing software that will execute quickly on a parallel computer. The second difficulty is transporting software between dissimilar parallel computers. In general, we expect that more hardware-specific information will be included in software designs for parallel computers than in designs for sequential computers. This inclusion is an instance of portability being sacrificed for high performance. New parallel computers are being introduced frequently. Trying to keep one's software on the current high performance hardware, a software developer almost continually faces yet another expensive software transportation. The problem of the proposed research is to create a design methodology that helps designers to more precisely control both portability and hardware-specific programming details. The proposed research emphasizes programming for scientific applications. We completed our study of the parallelizability of a subsystem of the NASA Earth Radiation Budget Experiment (ERBE) data processing system. This work is summarized in section two. A more detailed description is provided in Appendix A ('Programming Practices to Support Eventual Parallelism'). Mr. Chrisman, a graduate student, wrote and successfully defended a Ph.D. dissertation proposal which describes our research associated with the issues of software portability and high performance. The list of research tasks are specified in the proposal. The proposal 'A Design Methodology for Portable Software on Parallel Computers' is summarized in section three and is provided in its entirety in Appendix B. We are currently studying a proposed subsystem of the NASA Clouds and the Earth's Radiant Energy System (CERES) data processing system. This software is the proof-of-concept for the Ph.D. dissertation. We have implemented and measured the performance of a portion of this subsystem on the Intel iPSC/2 parallel computer. These results are provided in section four. Our future work is summarized in section five, our acknowledgements are stated in section six, and references for published papers associated with NAG-1-995 are provided in section seven.

Nicol, David M.↗

Laboratory Testing of Portable Air Cleaner Products for Energy Efficiency and Clean Air Performance at Various Fan Settings [Slides]

This document contains the results of a product review and laboratory testing experiment of portable air cleaning devices (PACs). Thirty-four commercially available PACs were reviewed, and from those, 7 products were purchased and underwent AHAM AC-7 clean air delivery rate (CADR) testing. The following research questions were explored: 1) What is the relationship between clean air delivery rate (CADR) and energy across a range of products and what is the general decrement in CADR as a result of operating at lower fan speeds? 2) How might the multiple units at lower power compare to fewer, larger units at full power in terms of energy efficiency and total cost? 3) How do the measured CADR, power, and efficiency compare to the manufacturer specs? 4) Are there product attributes that can be identified that might contribute to better CADR/W performance?

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Exploring code portability solutions for HEP with a particle tracking test code

Traditionally, high energy physics (HEP) experiments have relied on x86 CPUs for the majority of their significant computing needs. As the field looks ahead to the next generation of experiments such as DUNE and the High-Luminosity LHC, the computing demands are expected to increase dramatically. To cope with this increase, it will be necessary to take advantage of all available computing resources, including GPUs from different vendors. A broad landscape of code portability tools—including compiler pragma-based approaches, abstraction libraries, and other tools—allow the same source code to run efficiently on multiple architectures. In this paper, we use a test code taken from a HEP tracking algorithm to compare the performance and experience of implementing different portability solutions. While in several cases portable implementations perform close to the reference code version, we find that the performance varies significantly depending on the details of the implementation. Achieving optimal performance is not easy, even for relatively simple applications such as the test codes considered in this work. Several factors can affect the performance, such as the choice of the memory layout, the memory pinning strategy, and the compiler used. The compilers and tools are being actively developed, so future developments may be critical for their deployment in HEP experiments.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

A portable battery for objective, non-obstrusive measures of human performances

The need for a standardized battery of human performance tests to measure the effects of various treatments is pointed out. Progress in such a program is reported. Three batteries are available which differ in length and the number of tests in the battery. All tests are implemented on a portable, lap held, briefcase size microprocessor. Performances measured include: information processing, memory, visual perception, reasoning, and motor skills, programs to determine norms, reliabilities, stabilities, factor structure of tests, comparisons with marker tests, apparatus suitability. Rationale for the battery is provided.

Kennedy, R. S.↗

The Minos Computing Library: Efficient Parallel Programming for Extremely Heterogeneous Systems

Hardware specialization has become the silver bullet to achieve efficient high performance, from Systems-on-Chip systems, where hardware specialization can be ``extreme'', to large-scale HPC systems. As the complexity of the systems increases, so does the complexity of programming such architectures in a portable way. This work introduces the Minos Computing Library (MCL), as system software, programming model, and programming model runtime that facilitate programming extremely heterogeneous systems. MCL supports the execution of several multi-threaded applications within the same compute node, performs asynchronous execution of application tasks, efficiently balances computation across hardware resources, and provides performance portability. We show that code developed on a personal desktop automatically scales up to fully utilize powerful workstations with 8 GPUs and down to power-efficient embedded systems. MCL provides up to 17.5x speedup over OpenCL on NVIDIA DGX-1 systems and up to 1.88x speedup on single-GPU systems. In multi-application workloads, MCL dynamically resource allocation provides up to 2.43x performance improvement over manual, static allocation of computing resources.

Gioiosa, Roberto↗