Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “portable”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

Portable fiber optic sensor for rare earth elements and other critical metals using photoluminescence methods

Rare earth elements and other metals are vital to a range of technologies that are used in the energy and defense sectors. However, monopolistic market conditions have caused significant concern over the stability of the critical metal supply chain, and this has spurred extensive efforts in many nations to produce these metals domestically, both from conventional sources such as mining and well as from unconventional sources such as coal and its utilization byproducts. Slow and expensive characterization methods pose a significant barrier for both metals prospecting and process monitoring. A promising solution to this challenge is the development of highly sensitive luminescent sensors for metals, which can offer low costs, portability, and sensitivity. Anionic zinc adeninate metal-organic frameworks (BioMOFs) are known to distinguishing and detect part-per-billion levels of terbium, europium, samarium, and dysprosium in water by sensitizing the narrow, element-specific emission bands from these lanthanides. Here, a BioMOF material is immobilized onto a large diameter, solarization-resistant fiber optic tip integrated with a portable, low-cost spectrometer for rare earth element sensing. Immobilizing the sensing material on fiber instead of dispersing the sensing material in solution offers several advantages: it facilitates solvent removal, which enhances luminescent signal from the sensitized lanthanides, and it also allows the BioMOF to be recycled for multiple uses. The sensing system was deployed on a simulated process stream and exhibited qualitative agreement with inductively-coupled plasma mass spectrometry for terbium and europium detection, highlighting the potential for the sensing system to be deployed for real-world applications. By using different sensing materials, the same portable sensor may be deployed to detect other energy relevant metals such as cobalt, providing a cost-effective and sensitive platform for critical metal characterization.

Crawford, Scott↗

Oak Ridge National Laboratory Modernizing the Kokkos Build System: Using CMake to Encapsulate the Complexity of Build Instructions for Performance Portable Libraries

Kokkos, a C++ library focused on performance portability, requires a build system that can work with a variety of compilers and hardware. Ideally, users need only select the compiler and architecture and should not have to know or specify how programs using Kokkos are built. CMake can be used to create a flexible, robust build system and automatically configures compilers and settings based on the user’s inputs. Nevertheless, Kokkos’ requirements as a performance portability library for the build system exceed CMake’s current capabilities. This report describes the requirements, solutions, and testing of various implementations to create a CMake-based build system suitable for Kokkos. It compares the strengths and shortcomings of the approaches and evaluates the implementations with respect to the requirements. Because no solution was found to meet all of the requirements, the Kokkos team engaged with the CMake development team to discuss and plan a path toward support for performance-portable build systems in CMake in the future.

97 MATHEMATICS AND COMPUTING↗

Low-Cost and Portable Biosensor Based on Monitoring Impedance Changes in Aptamer-Functionalized Nanoporous Anodized Aluminum Oxide Membrane

We report a low-cost, portable biosensor composed of an aptamer-functionalized nanoporous anodic aluminum oxide (NAAO) membrane and a commercial microcontroller chip-based impedance reader suitable for electrochemical impedance spectroscopy (EIS)-based sensing. The biosensor consists of two chambers separated by an aptamer-functionalized NAAO membrane, and the impedance reader is utilized to monitor transmembrane impedance changes. The biosensor is utilized to detect amodiaquine molecules using an amodiaquine-binding aptamer (OR7)-functionalized membrane. The aptamer-functionalized membrane is exposed to different concentrations of amodiaquine molecules to characterize the sensitivity of the sensor response. The specificity of the sensor response is characterized by exposure to varying concentrations of chloroquine, which is similar in structure to amodiaquine but does not bind to the OR7 aptamer. A commercial potentiostat is also used to measure the sensor response for amodiaquine and chloroquine. The sensing response measured using both the portable impedance reader and the commercial potentiostat showed a similar dynamic response and detection threshold. The specific and sensitive sensing results for amodiaquine demonstrate the efficacy of the low-cost and portable biosensor.

60 APPLIED LIFE SCIENCES↗

TChem-atm (v2.0.0): scalable performance-portable multiphase atmospheric chemistry

We present TChem-atm, a performance-portable approach that enables efficient simulation of chemically detailed and multiphase atmospheric chemistry on modern heterogeneous computing architectures. Unlike previous efforts that rely on architecture-specific code or focus exclusively on gas-phase chemistry, TChem-atm supports fully coupled gas–aerosol systems with execution across CPUs, NVIDIA GPUs, and AMD GPUs through the Kokkos programming model. It integrates the flexible multiphase capabilities of the Community Atmospheric Model Chemistry Package (CAMP) with the high-performance kinetic routines of TChem, and includes automatic Jacobian construction with support for a range of stiff ODE solvers. In a proof-of-concept integration with the particle-resolved model PartMC, TChem-atm reproduces the existing PartMC–CAMP implementation within solver tolerances and delivers substantial GPU speedups, especially for large particle populations. Performance benchmarks reveal substantial speedups on GPU platforms, particularly for large particle populations, with consistent results across hardware backends. TChem-atm enables performance-portable execution across CPUs and GPUs, though optimal efficiency may require modest architecture-specific tuning (e.g., team and vector sizes), with up to a twofold improvement on the NVIDIA H100. It directly supports sectional and particle-resolved host models, while modal aerosol schemes require minor adaptation to provide particle-scale quantities such as representative diameters. By enabling chemically detailed, multiphase simulations with performance portability and host-model flexibility, TChem-atm facilitates the incorporation of advanced chemistry into atmospheric models.

Díaz-Ibarra, Oscar Homero [Sandia National Laborat↗

Portable generator having a configurable load bank

A portable generator includes a combustion engine. The portable generator includes an electric generator coupled to the combustion engine. The portable generator can include a load bank. When the electric generator operates at a first voltage and generates less than a threshold amount, the load bank is coupled to the electric generator in a first configuration. When the electric generator operates at a second voltage that is different than the first voltage, the load bank is coupled to the electric generator in a second configuration that is different than the first configuration.

42 ENGINEERING↗

Portability: A Necessary Approach for Future Scientific Software

Today's world of scientific software for High Energy Physics (HEP) is powered by x86 code, while the future will be much more reliant on accelerators like GPUs and FPGAs. The portable parallelization strategies (PPS) project of the High Energy Physics Center for Computational Excellence (HEP/CCE) is investigating solutions for portability techniques that will allow the coding of an algorithm once, and the ability to execute it on a variety of hardware products from many vendors, especially including accelerators. We think without these solutions, the scientific success of our experiments and endeavors is in danger, as software development could be expert driven and costly to be able to run on available hardware infrastructure. We think the best solution for the community would be an extension to the C++ standard with a very low entry bar for users, supporting all hardware forms and vendors. We are very far from that ideal though. We argue that in the future, as a community, we need to request and work on portability solutions and strive to reach this ideal.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Evaluating Portable Parallelization Strategies for Heterogeneous Architectures in High Energy Physics

High-energy physics (HEP) experiments have developed millions of lines of code over decades that are optimized to run on traditional x86 CPU systems. However, we are seeing a rapidly increasing fraction of floating point computing power in leadership-class computing facilities and traditional data centers coming from new accelerator architectures, such as GPUs. HEP experiments are now faced with the untenable prospect of rewriting millions of lines of x86 CPU code, for the increasingly dominant architectures found in these computational accelerators. This task is made more challenging by the architecture-specific languages and APIs promoted by manufacturers such as NVIDIA, Intel and AMD. Producing multiple, architecture-specific implementations is not a viable scenario, given the available person power and code maintenance issues. The Portable Parallelization Strategies team of the HEP Center for Computational Excellence is investigating the use of Kokkos, SYCL, OpenMP, std::execution::parallel and alpaka as potential portability solutions that promise to execute on multiple architectures from the same source code, using representative use cases from major HEP experiments, including the DUNE experiment of the Long Baseline Neutrino Facility, and the ATLAS and CMS experiments of the Large Hadron Collider. This cross-cutting evaluation of portability solutions using real applications will help inform and guide the HEP community when choosing their software and hardware suites for the next generation of experimental frameworks. We present the outcomes of our studies, including performance metrics, porting challenges, API evaluations, and build system integration.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Application of performance portability solutions for GPUs and many-core CPUs to track reconstruction kernels

Next generation High-Energy Physics (HEP) experiments are presented with significant computational challenges, both in terms of data volume and processing power. Using compute accelerators, such as GPUs, is one of the promising ways to provide the necessary computational power to meet the challenge. The current programming models for compute accelerators often involve using architecture-specific programming languages promoted by the hardware vendors and hence limit the set of platforms that the code can run on. Developing software with platform restrictions is especially unfeasible for HEP communities as it takes significant effort to convert typical HEP algorithms into ones that are efficient for compute accelerators. Multiple performance portability solutions have recently emerged and provide an alternative path for using compute accelerators, which allow the code to be executed on hardware from different vendors. We apply several portability solutions, such as Kokkos, SYCL, C++17 std::execution::par and Alpaka, on two mini-apps extracted from the mkFit project: p2z and p2r. These apps include basic kernels for a Kalman filter track fit, such as propagation and update of track parameters, for detectors at a fixed z or fixed r position, respectively. The two mini-apps explore different memory layout formats. We report on the development experience with different portability solutions, as well as their performance on GPUs and many-core CPUs, measured as the throughput of the kernels from different GPU and CPU vendors such as NVIDIA, AMD and Intel.

Kwok, Ka Martin↗

Scaling and performance portability of the particle-in-cell scheme for plasma physics applications through mini-apps targeting exascale architectures

We perform a scaling and performance portability study of the particle-in-cell scheme for plasma physics applications through a set of mini-apps we name "Alpine", which can make use of exascale computing capabilities. The mini-apps are based on Independent Parallel Particle Layer, a framework that is designed around performance portable and dimension independent particles and fields. We benchmark the simulations with varying parameters such as grid resolutions (5123 to 20483) and number of simulation particles (109 to 1011) with the following mini-apps: weak and strong Landau damping, bump-on-tail and two-stream instabilities, and the dynamics of an electron bunch in a charge-neutral Penning trap. We show strong and weak scaling and analyze the performance of different components on several pre-exascale architectures such as Piz-Daint, Cori, Summit and Perlmutter. While the scaling and portability study helps identify the performance critical components of the particle-in-cell scheme in the current state-of-the-art computing architectures, the mini-apps by themselves can be used to develop new algorithms and optimize their high performance implementations targeting exascale architectures.

Muralikrishnan, Sriramkrishnan↗

Ir-192 radioisotope replacement with a hand-portable 1 MeV Ku-band electron linear accelerator

Although linear accelerators are used in many security, industrial and medical applications, the existing technologies are too large and expensive for several critical applications such as radioactive source replacement, field radiography and mobile cargo scanners. Here, one of the main requirements for these sources is to be highly portable to allow field operation. In response to this problem, RadiaBeam has designed a hand-portable 1 MeV X-ray source, scalable to higher energies, based on Ku-band split electron linac, that can be used for Ir-192 radioisotope replacement. In this paper, we present its multiphysics and engineering design studies, as well as an accelerating structure prototype along with RF measurements.

43 PARTICLE ACCELERATORS↗

Implementing a neural network interatomic model with performance portability for emerging exascale architectures

The two main thrusts of computational science are increasingly accurate predictions and faster calculations; to this end, the zeitgeist in molecular dynamics (MD) simulations is pursuing machine learned and data driven interatomic models, e.g. neural network potentials, and novel hardware architectures, e.g. GPUs. Current implementations of neural network potentials are orders of magnitude slower than traditional interatomic models and while looming exascale computing offers the ability to run large, accurate simulations with these models, achieving portable performance for MD with new and varied exascale hardware requires rethinking traditional algorithms, using novel data structures, and library solutions. We re-implement a neural network interatomic model in CabanaMD, an MD proxy application, built on libraries developed for performance portability. Our implementation shows significantly improved thread scaling in this complex kernel as compared to a current LAMMPS implementation, across both strong and weak scaling. Our single-source solution enables simulations up to 20 million atoms on a single CPU node and 4 million atoms with improved performance on a single GPU. Furthermore, we also explore parallelism and data layout choices (using flexible data structures called AoSoAs) and their effect on performance, seeing up to ~50% and ~5% improvements in performance on a GPU by choosing the right level of parallelism and data layout respectively.

97 MATHEMATICS AND COMPUTING↗

Achieving performance portability in Gaussian basis set density functional theory on accelerator based architectures in NWChemEx

The numerical integration of the exchange–correlation (XC) potential is one of the primary computational bottlenecks in Gaussian basis set Kohn–Sham density functional theory (KS-DFT). To achieve optimal performance and accuracy, care must be taken in this numerical integration to preserve local sparsity as to allow for near linear weak scaling with system size. This leads to an integration scheme with several performance critical kernels which must be hand optimized for each architecture of interest. As the set of available accelerator hardware goes more diverse, a key challenge for developers of KS-DFT software is to maintain performance portability across a wide range of computational architectures. In this article, we examine a modular software design pattern which decouples the implementation details of performance critical kernels from the expression of high-level algorithmic workflows in a device-agnostic language such as C++; thus allowing for developers to target existing and emerging accelerator hardware within a single code base. We consider the efficacy of such a design pattern in the numerical integration of the XC potential by demonstrating its ability to achieve performance portability across a set of accelerator architectures which are representative of those on current and future U.S. Department of Energy Leadership Computing Facilities.

97 MATHEMATICS AND COMPUTING↗

Toward performance-portable PETSc for GPU-based exascale systems

The Portable Extensible Toolkit for Scientific computation (PETSc) library delivers scalable solvers for nonlinear time-dependent differential and algebraic equations and for numerical optimization. The PETSc design for performance portability addresses fundamental GPU accelerator challenges and stresses flexibility and extensibility by separating the programming model used by the application from that used by the library, and it enables application developers to use their preferred programming model, such as Kokkos, RAJA, SYCL, HIP, CUDA, or OpenCL, on upcoming exascale systems. Furthermore, a blueprint for using GPUs from PETSc-based codes is provided, and case studies emphasize the flexibility and high performance achieved on current GPU-based systems.

97 MATHEMATICS AND COMPUTING↗

Ku-band electron linac for battery-powered hand-portable 2-MeV X-ray generator

X-ray generators, producing radiation in MeV range, are a critical tool for radiography, non-destructive testing and security applications. Field operation of such source requires them to be hand-portable, autonomous and allow parameter adjustability. RF linear accelerators can serve as a flexible, reliable, and robust radiation generator alternative to dangerous radioisotopes and bulky betatrons that are currently used for field radiography if their size, weight, cost, and imaging performance are matched to these sources. Here, in this paper, we present the design and test results of a 2 MeV Ku-band electron linac for a hand-portable X-ray generator system for field radiography being developed by RadiaBeam. The dramatic scale of miniaturization and cost-reduction is achieved thanks to the implementation of innovative technologies such as air-cooled Ku-band air-traffic control magnetrons, a split accelerating structure fabrication technique, and solid-state Marx modulators. This paper presents the design of the first prototype of the accelerator, its operation from Li-Ion batteries, as well as high-power and beam measurements.

43 PARTICLE ACCELERATORS↗

A portable and reusable sensor system based on graphene for real-time and sensitive detection of lead ions in water

Long-term exposure to Pb 2+ can cause irreversible damage to the nervous, cardiovascular, and reproductive systems. Therefore, developing a fast and sensitive detection system capable of monitoring minuscule concentrations of Pb 2+ is essential. In this study, we demonstrated a fully portable sensor system enabling rapid, sensitive, and real-time monitoring of Pb 2+ . The sensor system adopted the remote-gate field-effect transistor (RGFET) detection scheme and was easy to operate, even for non-experts. The sensor system comprised two printed circuit boards (PCBs): a sensor PCB with a remote-gate electrode and an analyzer PCB with a metal-oxide-semiconductor field-effect transistor (MOSFET) transducer and peripheral electronics to manage sensor signals. To achieve a high sensitivity for Pb 2+ , we utilized graphene ink drop-casted on the sensor PCB as a sensing membrane. The graphene film was easy to deposit and remove, enabling the sensor PCB to be reused multiple times. The sensor system was further linked to a smartphone application that instantly monitors the sensor response, allowing for rapid point-of-use detection. The sensor exhibited a high sensitivity of 21.7% when the limit of detection (LOD) value of 1 nM (∼0.2 ppb) was detected, and the typical detection time for each sample was approximately 60 seconds. This portable sensor system advances sensing technologies and could potentially supplement expensive, laborious conventional sensing equipment.

54 ENVIRONMENTAL SCIENCES↗

Ultralightweight Power System for Human-Portable Linac-Based X-Ray Sources

Industrial human-portable X-ray sources are widely used by security, nuclear safeguard, and defense agencies. However, the employed sources have significant energy, dose, size, weight, and power (SWaP) limitations, greatly affecting their practical application. RF linear accelerators (linacs) can serve as a flexible, reliable, and robust type of X-ray source if they can match the size, weight, cost, and imaging performance requirements of conventional ones. One of the most critical elements affecting these parameters is the high-voltage pulsed power supply system or modulator, which can make the largest contribution to the total weight and dimensions of the accelerator. Here, in this article, we present the design and demonstration results of a novel ultra lightweight power system based on a 24-kV solid-state Marx modulator for a hand-portable 0.15–2.0-MeV Ku -band linac-based X-ray source.

47 OTHER INSTRUMENTATION↗

ArborX: A Performance Portable Geometric Search Library

Searching for geometric objects that are close in space is a fundamental component of many applications. The performance of search algorithms comes to the forefront as the size of a problem increases both in terms of total object count as well as in the total number of search queries performed. Scientific applications requiring modern leadership-class supercomputers also pose an additional requirement of performance portability, i.e., being able to efficiently utilize a variety of hardware architectures. In this article, we introduce a new open-source C++ search library, ArborX, which we have designed for modern supercomputing architectures. Herein, we examine scalable search algorithms with a focus on performance, including a highly efficient parallel bounding volume hierarchy implementation, and propose a flexible interface making it easy to integrate with existing applications. We demonstrate the performance portability of ArborX on multi-core CPUs and GPUs and compare it to the state-of-the-art libraries such as Boost.Geometry.Index and nanoflann.

97 MATHEMATICS AND COMPUTING↗

SeeQ: A Programming Model for Portable Data-Driven Building Applications

This paper introduces SeeQ, a programming model and an abstraction framework that facilitates the development of portable data- driven building applications. Data-driven approaches can provide insights into building operations and guide decision-making to achieve operational objectives. Yet the configuration of such applications per building requires extensive effort and tacit knowledge. In SeeQ, we propose a portable programming model and build a software system that enables self-configuration and execution across diverse buildings. The configuration of each building is captured in a unified data model - in this paper, we work with the Brick ontology without loss of generality. SeeQ focuses on the distinction between the application logic and the configuration of an application against building-specific data inputs and systems. We test the proposed approach by configuring and deploying a diverse range of applications across five heterogeneous real-world buildings. The analysis shows the potential of SeeQ to significantly reduce the efforts associated with the delivery of building analytics.

analytics↗