Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “offloading”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

305 records · Page 17

Volumetric Assessment of UPRITE Exercises From Marker-Based Motion Capture

BACKGROUND Lack of volumetric data on full-body movement of exercises presents a challenge to ensuring the fit of crew member’s full range of motion on the International Space Station (ISS). The Upright Proprioception Retention via In-flight Training and Evaluation (UPRITE) is a sensorimotor countermeasure device designed for maintaining crew members’ proprioception in a microgravity environment. A footplate—attached to a static base—rotates in two degrees of freedom (pitch and roll) up to a 20 deg angle. An initial volumetric assessment assuming an upright standing posture produced a cone-like shape with a narrow bottom and wide top. Such general volumetric assessments risk creating an overly conservative volume estimate, taking up more space than is necessary on the already limited interior space of the ISS, and neglecting necessary volume due to oversimplifying assumptions. Rather, higher-fidelity volumetric assessments offer more comprehensive insights in an environment where every area counts. The main objective of this work is to provide the spatial parameters of exercises on the UPRITE such that it is placed on the ISS according to its volumetric demands or that usage is adjusted to fit the available space. METHODS In 2023, a data collection was performed originally to inform loads and dynamics of system use and was recently leveraged for volumetric assessment. Three human subjects representing different body types (~63-76 inches in stature) performed a variety of board manipulations using UPRITE with body weight offload. The test collected the 3D positional data of a modified full-body Plug-in Gait marker set [1] via a 16-camera OptiTrack MoCap system. After processing – filling marker gaps and trimming data – in OptiTrack Motive, the recorded marker location data, which included device markers, was exported to a readable trajectory file. To accurately represent the full volume defining landmarks, additional markers were digitally added to an unscaled Modified Full Body Model [2]. The model was then scaled according to its subject parameters upon which an inverse kinematics analysis was performed. A custom plugin yielded model marker location data files. Volumetric analyses were performed on the recorded trajectory and model trajectory files using a custom Python-built tool that extracted the marker location data and plotted it in a 3D space. Concerned with only the maximum volume of the motion, a 3D convex hull analysis was applied to the plot, extracting the vertices or external points of the eventual 3D CAD output, dubbed aptly as a “volume shell”. This overall approach was based on guidance in a NASA-STD-3001 Technical Brief [3]. RESULTS AND DISCUSSION Batch volumetric assessment on the exercises for each subject was performed, producing high-fidelity volume shells in minimal time. Preliminary results highlighted the value in higher-fidelity volumes based on collected data when possible. For example, revolving a single posture in the cone assessment would not have sufficiently captured a single leg stance; rather, it would need to involve swinging the leg both forward and back. Additional observations and the maximal dimensions of the volumes, including those based on scaled data for ISS anthropometric requirements, will be presented at the Human Research Program Investigator’s Workshop. CONCLUSIONS While this work’s primary objective was for the UPRITE-to-ISS integration, the tool built to conduct this analysis has wide applications for future exercise systems as an informational tool for optimal device placement. The tool and its findings also have implications for exercise device design and spacecraft interior considerations on Gateway, the Lunar Pressurized Rover, and beyond. REFERENCES [1] Bell, C. A., et al. (2023) Recent Improvements and Verification of a Full Body Model in OpenSim. NASA Human Research Program Investigator’s Workshop. https://ntrs.nasa.gov/citations/20230001080 [2] Lostroscio, K., et al (2023) The Digital Astronaut Simulation. AHFE International Conference on Human Factors in Design, Engineering, and Computing for All. [3] Exercise Overview. (2023) NASA-STD-3001 Technical Brief. https://www.nasa.gov/wp-content/uploads/2023/12/ochmo-tb-031-exercise-overview.pdf?emrc=9d454c?emrc=9d454c

L D Quinto↗

Using DIC for Long Slender Structures

High-strain composite deployable structures have been developed for systems such as solar arrays, camera masts or solar sailing propulsion elements. Composite booms in such applications are often flattened and then rolled into a small footprint for low-packaged volume and are then deployed in space. There is a need to obtain deformation for long, slender composite booms on earth through gravity offloading by suspending them vertically and applying distal end (tip) loads. Three-dimensional digital image correlation (3D-DIC), along with other measurement techniques, were used to obtain strain and displacement along the length of a 7.5 m subscale composite Triangular, Rollable, and Collapsible (TRAC) boom in preparation for full scale testing of a 30 m boom. However, incorporating 3D-DIC as a primary measurement tool on long, slender, high-aspect-ratio boom structures presents significant challenges. Challenges include small correlated area due to high aspect ratio, limited standoff distance due to size of test area, coordinate system alignment of multiple camera systems along the length of the boom, nodal mesh extraction for adequate test/analysis correlation, as well as measurement comparison between DIC and other instrumentation used such as fiber optic strain sensing (FOSS) and laser displacement tracking. The contents of the proposed paper will focus on techniques and methods for overcoming the previously mentioned challenges associated with applying 3D-DIC to long, slender boom structures. Results from subscale test along with lessons learned will be discussed.

Deployable Boom↗

CIPHER: Egress Fitness

The transition between gravity environments will involve one of the most complex, high-risk phases of exploration missions. The reduced functional capacity caused by physiological deconditioning adaptations in microgravity coupled with the stressors of re-entry into partial gravity environments will increase risks to crew, even with rigorous adherence to inflight countermeasures. Specifically, two high-risk scenarios may be required to be performed soon after gravity transitions: 1) nominal and/or emergency unassisted capsule egress task after return to Earth, and 2) planetary extravehicular activity (EVA) soon after landing on Mars or the Moon. Quantification of crewmember’s functional performance after long-duration spaceflight is necessary to inform concepts of operations for future exploration missions. The overarching aim of this study is to quantify post-landing functional performance with deconditioning after long-duration ISS missions. This study is broken down into two phases. Phase 1 includes a pilot study to assess the overall feasibility and demonstrate the capability to perform mission-like tasks shortly after landing. Phase 2, the Egress Fitness study, which is part of the Complement of Integrated Protocols for Human Exploration Research (CIPHER), uses a task-based approach to characterize functional performance in long-duration ISS crewmembers before flight and shortly after return to Earth. The pilot and full Egress Fitness study includes pre-flight and post-flight testing of simulated emergency egress out of a functional capsule mockup and a Mars gravity EVA simulation at the Active Response Gravity Offload System (ARGOS) facility. The EVA simulation tasks include suit donning, hatch egress, ladder descent, task board cable operations, baggage transfer over sand/rocky regolith, alignment with a rear entry port, and suit egress. The post-flight simulated capsule egress test occurs 1–4 h after landing and the planetary EVA simulation occurs 18–36 h after landing. The full CIPHER Egress Fitness study has additional pre-flight sessions, longer EVA tasks that include traverse and geology sampling, and post-flight sessions on R+1, 4, and 8 to characterize the timeframe of recovery. Pilot Egress Fitness has completed baseline and post-flight testing on four crewmembers. That study remains open. Originally this was to cover the Boeing CFT mission, but now also includes private astronauts on commercial spaceflights. CIPHER study data collection is ongoing with 2 subjects completed and 4 additional subjects consented. This study will quantify post-landing functional performance in operationally relevant simulations to help inform fitness for duty standards and future planetary concepts of operations shortly after landing.

Jason Norcross↗

CCAMP: An Integrated Translation and Optimization Framework for OpenACC and OpenMP

Heterogeneous computing and exploration into specialized accelerators are inevitable in current and future supercomputers. Although this diversity of devices is promising for performance, the array of architectures presents programming challenges. High-level programming strategies have emerged to face these challenges, such as the OpenMP offloading model and OpenACC. The varying levels of support for these standards, however, within vendor-specific and open-source tools, as well as the lack of performance portability across devices, have prevented the standards from achieving their goals. To address these shortcomings, we present CCAMP, an OpenMP and OpenACC interoperable framework. CCAMP provides two primary facilities: language translation between the two standards and device-specific directive optimization within each standard. We show that by using the CCAMP framework, programmers can easily transplant non-portable code into new ecosystems for new architectures. Additionally, by using CCAMP device-specific directive optimizations, users can achieve optimized performance across architectures using a single source code.

Lambert, Jacob↗

C-SAW: a framework for graph sampling and random walk on GPUs

Many applications require to learn, mine, analyze and visualize large-scale graphs. These graphs are often too large to be addressed efficiently using conventional graph processing technologies. Fortunately, recent research efforts find out graph sampling and random walk, which significantly reduce the size of original graphs, can benefit the tasks of learning, mining, analyzing and visualizing large graphs by capturing the desirable graph properties. This paper introduces C-SAW, the first framework that accelerates Sampling and Random Walk framework on GPUs. Particularly, C-SAW makes three contributions: First, our framework provides a generic API which allows users to implement a wide range of sampling and random walk algorithms with ease. Second, offloading this framework on GPU, we introduce warp-centric parallel selection, and two novel optimizations for collision migration. Third, towards supporting graphs that exceed the GPU memory capacity, we introduce efficient data transfer optimizations for out-of-memory and multi-GPU sampling, such as workload-aware scheduling and batched multi-instance sampling. Taken together, our framework constantly outperforms the state of the art projects in addition to the capability of supporting a wide range of sampling and random walk algorithms.

97 MATHEMATICS AND COMPUTING↗

An Inner-Loop Control Method for the Filter-less, Voltage Sensor-less, and PLL-less Grid-Following Inverter-Based Resource

This paper presents a novel inner-loop control method for the inverter-based resource (IBR). The innovative concepts include removing the voltage sensors at the point of common coupling (PCC), removing the inverter interface inductance, and removing the traditional phase-locked loop (PLL) circuits. Simulation is conducted to verify the feasibility of the method. Furthermore, the virtual impedance is applied to the control loop to improve the current THD and dynamics. Compared to the traditional control, the proposed one helps offload the system by reducing the bulky inductors and voltage sensors without compromising the control performance.

grid-forming inverter, inner-loop control, filterl↗

SPEChpc 2021 Benchmark Suites for Modern HPC Systems

The SPEChpc 2021 suites are application-based benchmarks de- signed to measure performance of modern HPC systems. The bench- marks support MPI, MPI+OpenMP, MPI+OpenMP target offload, MPI+OpenACC and are portable across all major HPC platforms.

Boehm, Swen↗

Union: A Unified HW-SW Co-Design Ecosystem in MLIR for Evaluating Tensor Operationson Spatial Accelerators

To meet the extreme compute demands for deep learning across commercial and scientific applications, dataflow accelerators are becoming increasingly popular. While these“domain-specific” accelerators are not fully programmable like CPUs and GPUs, they retain varying levels of flexibility with respect to data orchestration, i.e., dataflow and tiling optimizations to enhance efficiency. There are several challenges when designing new algorithms and mapping approaches to execute the algorithms for a target problem on new hardware. Previous works have addressed these challenges individually. To address this challenge as a whole, in this work, we present an HW-SW co-design ecosystem for spatial accelerators called Union within the popular MLIR compiler infrastructure. Our framework allows exploring different algorithms and their mappings on several accelerator cost models. Union also includes a plug-and-play library of accelerator cost models and mappers which can easily be extended. The algorithms and accelerator cost models are connected via a novel mapping abstraction that captures the map space of spatial accelerators which can be systematically pruned based on constraints from the hardware, workload, and mapper. We demonstrate the value of Union for the community with several case studies which examine offloading different tensor operations (CONV/GEMM/Tensor Contraction) on diverse accelerator architectures using different mapping schemes.

Jeong, Geonhwa↗

symPACK: A GPU-Capable Fan-Out Sparse Cholesky Solver

Sparse symmetric positive definite systems of equations are ubiquitous in scientific workloads and applications. Parallel sparse Cholesky factorization is the method of choice for solving such linear systems. Therefore, the development of parallel sparse Cholesky codes that can efficiently run on today’s large-scale heterogeneous distributed-memory platforms is of vital importance. Modern supercomputers offer nodes that contain a mix of CPUs and GPUs. To fully utilize the computing power of these nodes, scientific codes must be adapted to offload expensive computations to GPUs. We present symPACK, a GPU-capable parallel sparse Cholesky solver that uses one-sided communication primitives and remote procedure calls provided by the UPC++ library. We also utilize the UPC++ "memory kinds" feature to enable efficient communication of GPU-resident data. We show that on a number of large problems, symPACK outperforms comparable state-of-the-art GPU-capable Cholesky factorization codes by up to 14x on the NERSC Perlmutter supercomputer.

Bellavita, Julian↗

Methods, systems, and apparatuses for calculating global fluence for neutron and photon monte carlo transport using expected value estimators

Global fluence estimators may be calculated on accelerators and processors for neutron and photon Monte Carlo transport. Monte Carlo random walk simulation may be performed on the processors and the calculation of a Volumetric-Ray-Casting (VRC) estimator may be offloaded to the accelerators. The VRC estimator may modify an expected-value estimator to extend a pseudo-particle ray along the direction of the emitted particle from source and collision event through not only the event volume, but also through all volumes that describe the problem geometry. Additionally, many pseudo-particle rays may be sampled per event, rather than just a single pseudo-particle ray per event, in order to provide more complete angular coverage.

Sweezy, Jeremy Ed↗

Integrating Artificial Intelligence into Science Gateways

Science gateways are altering the manner in which people interact with high performance computing (HPC) by providing a web browser based interface to advanced computing platforms. In particular, science gateways lower the barrier to using HPC by simplifying the process of submitting workloads to such systems and by offloading the efforts required to use HPC to the maintainers of the system. While science gateways decrease the time-to-science that comes with using such advanced systems, progress can still be made in improving the user's experience. In this paper we explore two strategies for integrating artificial intelligence tools commonly found in non-HPC service workflows: voice activated assistants and chatbots. Since August 2021, the HPC group at Idaho National Laboratory answers an average of 581 support tickets per month of which a large percentage could be addressed via these two strategies. This work defines the key capabilities that an HPC voice activated assistant and chatbot would need to address for a userbase consisting of largely non-expert users as well as a design for integration into the Open OnDemand science gateway.

97 MATHEMATICS AND COMPUTING↗

Accelerating detector simulations with Celeritas: profiling and performance optimizations

Celeritas is a GPU-optimized MC particle transport code designed to meet the growing computational demands of next-generation HEP experiments. It provides efficient simulation of EM physics processes in complex geometries with magnetic fields, detector hit scoring, and seamless integration into Geant4-driven applications to offload EM physics to GPUs. Recent efforts have focused on performance optimizations and expanding profiling capabilities. This paper presents some key advancements, including the integration of the Perfetto system profiling tool for detailed performance analysis and the development of track-sorting methods to improve computational efficiency.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Assessment of Cloud-based Applications for Enabling a Scalable Riskinformed Predictive Maintenance Strategy

The current light-water reactor fleet uses time-based maintenance strategies to achieve high-capacity factors. But to make nuclear more competitive in the energy market, these reactors could utilize emerging artificial intelligence (AI) and cloud computing technologies to achieve a cost-effective, predictive-maintenance strategy. This paper presents discussion and results on the application of cloud computing in the nuclear industry. The technical viability of cloud computing was analyzed using data from a boiling-water reactor’s safety relief valve. The models were hosted on three different systems: a local personal computer, Idaho National Laboratory’s high-performance computer system, and Microsoft Azure. The data were loaded and processed, and two types of models were trained in an A/B fashion. Based on the speed at which these actions were completed, it was determined that cloud computing affords adequate computing resources. Additionally, the computing power can scale with the demanded load. To enable cloud computing in the existing fleet, additional sensors, networks, and other requirements must be implemented to ensure a smooth transition from current maintenance strategies. However, the benefit is that the plants no longer need to manage their own servers, software, cybersecurity, and information technology support staff for in-house data analytics purpose. Many of these features can be offloaded to the cloud provider for a potential cost savings. Demonstrating how AI can improve the maintenance and operation of non-safety-related systems seems the likely path forward for implementing AI and cloud computing resources inside nuclear power plants.

azure↗

GenASiS Basics: Object-oriented utilitarian functionality for large-scale physics simulations (Version 4)

GenASiS Basics provides modern Fortran classes furnishing extensible object-oriented utilitarian functionality for large-scale physics simulations on distributed memory supercomputers. This functionality includes physical units and constants; display to the screen or standard output device; message passing; I/O to disk; and runtime parameter management and usage statistics. Herein, this revision—Version 4 of Basics—includes a name change and additions to functionality, including the facilitation of direct communication between GPUs.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Verification and Validation of a Conceptual Model of the Auto-Rigging Payload Handling and Off-Loading System Using LEGO Technic System and Three-Dimensional Printed Parts

A conceptual model (CM) can be used to validate a concept in modeling and simulation life cycles. During the 2020-2022 Coronavirus disease 2019 (COVID-19) pandemic, for employee safety NASA implemented center closures and mandatory telework for the entire workforce. During this challenging time, engineers and researchers at NASA Langley Research Center (LaRC) looked for safe and innovative approaches and methods to continue the development of CMs for various projects. Engineers and researchers at LaRC researched Auto-Rigging Payload Handling and Off-Loading System (ARPHOLS) for payload handing and off-loading a system on an inclined lunar lander deck. In this paper, the development of a CM and the verification and validation of a conceptual idea for ARPHOLS using a LEGO Technic system and three-dimensional printed parts is presented.

Payload offloading↗

Segmented Solid Surface Reflector Concentrically Stacked With Tubular Shape Memory Composite Hinges

A new architecture for solid surface reflector antennas scalable to sizes greater than 10 m is presented. The design uses compact, light, and simple advanced deployable structures to create sub-reflectors that can be assembled in space into larger units using a robotic arm. The seven-panel hexagonal sub-reflector is divided into hexagonal panels that stack concentrically and vertically. The central panel is connected to each side panel on the back side by a pair of tubular shape memory composite hinges that enable the required deployment kinematics with controlled dynamics. A secondary mechanism closes the interpanel gap. The focus of the paper is on the development of the sub-reflector elements, namely the tubular hinges that use embedded heaters and sensors for triggering and control, the actuation mechanisms, and the lightweight sandwich construction reflector panels. A parametric study using finite element analyses was conducted to assess how design features of the hinge affect its stowage and deployment dynamics. The preliminary component fabrication and testing results for the two-panel assembly breadboard model are outlined. Finally, the results of the test campaign with the brassboard reflector model are presented. A comparison of deployed reflector surface deviation between the measured surface after the stowage and deployment process and the pre-test scans and the nominal surface revealed root mean square errors of less than 1 mm, as required by X-band radiofrequency transmission.

gravity offloading↗

Segmented Hexagonal Antenna Reflector Concentrically Stacked Using Shape Memory Composite Tubular Hinges

A new architecture for solid surface reflector antennas scalable to sizes greater than 10 m is presented. The design uses compact, light, and simple advanced deployable structures to create sub-reflectors that can be assembled in space into larger units using a robotic arm. The seven-panel hexagonal sub-reflector is divided into hexagonal panels that stack concentrically and vertically. The central panel is connected to each side panel on the back side by a pair of tubular shape memory composite hinges that enable the required deployment kinematics with controlled dynamics. A secondary mechanism closes the interpanel gap. The focus of the paper is on the development of the sub-reflector elements, namely the tubular hinges that use embedded heaters and sensors for triggering and control, the actuation mechanisms, and the lightweight sandwich construction reflector panels. A parametric study using finite element analyses was conducted to assess how design features of the hinge affect its stowage and deployment dynamics. The preliminary component fabrication and testing results for the two-panel assembly breadboard model are outlined. Finally, the results of the test campaign with the brassboard reflector model are presented. A comparison of deployed reflector surface deviation between the measured surface after the stowage and deployment process and the pre-test scans and the nominal surface revealed root mean square errors of less than 1 mm, as required by X-band radiofrequency transmission.

gravity offloading methods↗