Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “parallelization”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 559 records · Page 31

OpenMDlr: parallel, open-source tools for general protein structure modeling and refinement from pairwise distances

Easy-to-use, open-source, general-purpose programs for modeling a protein structure from inter-atomic distances are needed for modeling from experimental data and refinement of predicted protein structures. OpenMDlr is an open-source Python package for modeling protein structures from pairwise distances between any atoms, and optionally, dihedral angles. Finally, we provide a user-friendly input format for harnessing modern biomolecular force fields in an easy-to-install package that can efficiently make use of multiple compute cores.

59 BASIC BIOLOGICAL SCIENCES↗

A mathematical model of asynchronous data flow in parallel computers *

Abstract We present a simplified model of data flow on processors in a high-performance computing framework involving computations necessitating inter-processor communications. From this ordinary differential model, we take its asymptotic limit, resulting in a model which treats the computer as a continuum of processors and data flow as an Eulerian fluid governed by a conservation law. We derive a Hamilton–Jacobi equation associated with this conservation law for which the existence and uniqueness of solutions can be proven. We then present the results of numerical experiments for both discrete and continuum models; these show a qualitative agreement between the two and the effect of variations in the computing environment’s processing capabilities on the progress of the modelled computation.

Barnard, Richard C.↗

Altermagnetic behavior in OsO 2 : Parallels with RuO 2

We investigate and compare the electronic structure of OsO 2 with the extensively studied RuO 2 . Calculations show that OsO 2 exhibits antiferromagnetism and spin splitting driven by crystal symmetry, a characteristic of altermagnetism and similar to RuO 2 . Examination of the Fermi surface, with and without spin-orbit coupling, reveals that OsO 2 has lower Fermi group velocities compared to RuO 2 , suggesting limited charge carrier mobility in OsO 2 . While this characteristic may constrain its performance in high-speed electronic transport applications, it may also enhance stability for spin-based information storage in spintronics. Additionally, comparison of the vibrational properties of these rutile oxide systems demonstrates dynamical stability, typical mass dependent behaviors, and favorable agreement with measured Raman data. The calculated phonon density of states for RuO 2 also agrees with our neutron scattering data. These observations substantiate the implications of the electronic and vibrational behaviors in OsO 2 and RuO 2 , encouraging further investigation of their potential for emerging technological applications.

36 MATERIALS SCIENCE↗

Parallel I/O Evaluation Techniques and Emerging HPC Workloads: A Perspective

Emerging workloads such as artificial intelligence, big data analytics and complex multi-step workflows alongside future exascale applications are anticipated future HPC workloads, which will result in a more diverse I/O system workload and even less predictable I/O behavior and access patterns. Along with the ever increasing gap between the compute and storage performance capabilities, the in-depth understanding of extreme-scale I/O behavior and the I/O performance modeling and prediction are essential tools of the large-scale I/O evaluation process for addressing the needs of extreme-scale hybrid workloads. In this survey article, we focus on the state-of-the-art of the I/O behavior and performance analysis process for HPC systems in a 5-year time window and identify future research challenges.

Neuwirth, Sarah↗

Modeling pre-Exascale AMR Parallel I/O Workloads via Proxy Applications

The present work investigates the modeling of preexascale input/output (I/O) workloads of Adaptive Mesh Refinement (AMR) simulations through a simple proxy application. We collect data from the AMReX Castro framework running on the Summit supercomputer for a wide range of scales and mesh partitions for the hydrodynamic Sedov case as a baseline to provide sufficient coverage to the formulated proxy model. The non-linear analysis data production rates are quantified as a function of a set of input parameters such as output frequency, grid size, number of levels, and the Courant-Friedrichs-Lewy (CFL) condition number for each rank, mesh level and simulation time step. Linear regression is then applied to formulate a simple analytical model which allows to translate AMReX inputs into MACSio proxy I/O application parameters, resulting in a simple “kernel” approximation for data production at each time step. Results show that MACSio can simulate actual AMReX nonlinear “static” I/O workloads to a certain degree of confidence on the Summit supercomputer using the present methodology. The goal is to provide an initial level of understanding of AMR I/O workloads via lightweight proxy applications models to facilitate autotune data management strategies in anticipation of exascale systems.

Godoy, William↗

A Parallel Machine Learning Workflow for Neutron Scattering Data Analysis

As part of a larger effort, this work-in-progress reports the possible advantages of modifying conventional workflows used to generate labelled training samples and train machine learning (ML) models on them. We compare results from three different workflows using neutron scattering data analysis as the motivating application and report about 20% improvement in speedup, with no appreciable loss of model accuracy, over a baseline workflow.

Wang, Tianle↗

Resilient Inverter-Driven Black Start with Collective Parallel Grid-Forming Operation

As modern power systems are experiencing exceptional changes with increasing penetrations of inverter-based resources (IBRs), system restoration using IBRs has received attention. Using local grid-forming (GFM) assets near consumers, engineered to establish grid voltages in the absence of a stiff grid, i.e., bottom-up restoration, a distribution system could obtain high system resilience by not relying on the bulk power system restoration, which requires significant human intervention and procedure. This paper studies the technical feasibility of the novel approach with detailed electromagnetic transient (EMT) simulations. To thoroughly evaluate the potential of GFM inverters and the technical challenges in IBR-driven black start, a detailed three-phase inverter model is developed, including negative-sequence control for voltage balance and a phase-by- phase current limiter to sustain momentary overloading during the black start. To examine dynamic aspects of the black-start process, the EMT simulation also models transformer and motor dynamics to emulate their inrush and startup behaviors as well as network dynamics. In addition, active involvement of grid- following distributed energy resources is also studied to facilitate the black-start process. By allowing multiple GFM inverters to collectively black start without leader-follower coordination, we demonstrate that a system can achieve high resilience even with a fraction of assets lost. Two test cases of inverter-driven black start, using two and one GFM inverters, respectively, for a heavily unbalanced 2-MVA distribution feeder are demonstrated. Takeaways for further study and field deployment are provided.

grid-forming inverter↗