Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “parallel systems”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 145 records · Page 8

Massively parallel, multiple-organ perfusion control system

A fluidic cartridge comprises a fluidic disk having a plurality of alignment openings; a fluidic chip comprising a body, one or more channels formed in the body in fluidic communications with input ports and output ports for transferring one or more fluids between the input ports and the output ports, and a plurality of protrusions formed on the body and received in the alignment openings of the fluidic disk for aligning the fluidic chip to the fluidic disk; an actuator operably engaging with the one or more channels for selectively and individually transferring the one or more fluids through the one or more channels from at least one of the input ports to at least one of the output ports at desired flow rates; and a tube member defining a cylindrical housing for accommodating the fluidic disk, the fluidic chip and the actuator therein.

Reiserer, Ronald S.↗

Update on Parallel Process Execution in the Next Generation System Analysis Model (NGSAM)

As of the end of 2022, it is estimated that over 90,000 metric tons of heavy metal (MTHM) of spent nuclear fuel (SNF) were stored at various commercial nuclear power reactor sites (both operating and shutdown) across the United States [1]. The Office of Storage and Transportation within the U.S. Department of Energy’s Office of Nuclear Energy is planning for the transportation, storage, and eventual disposal of SNF and high-level radioactive waste (HLW). To aid in this effort and inform decision-makers about the backend of the spent fuel cycle, systems analysis tools capable of analyzing the various options with respect to SNF and HLW management are being used as well as continuously improved to meet the evolving needs of the program. System analysts typically use these tools to vary underlying assumptions (shipping rates, available facilities, start dates, interim storage capacity, etc.) and study the associated system implications such as timing for clearing sites of SNF, various cost elements, transportation infrastructure acquisition needs, etc.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Design, Optimization, and Control of Floating Offshore Wind Farms for Optimal Energy Production (Final report)

The uncertainty and irregularity of ocean waves and the ocean environment is a major factor in the development of commercial scale floating wind turbines as the operation of floating structures in such an environment can lead to irregular and unpredictable loading, fatigue, and ultimately a reduction in the operational life of the turbine system which affects energy production over the lifetime of the turbine. Control solutions that can limit float motions and mitigate stressful events on the structure become essential for extending lifetime and limiting the operational uncertainty of a floating wind turbine. Digital twins are computational replicas of physical systems that operate in parallel with the operation of the physical system. Given advanced knowledge of a systems input, digital twins have the ability to predict the behavior of a system in advance, which can be valuable in the control of that system. In this project, we developed and assessed potential digital twin models developed in house and openly available (OpenFast) for use in the real time control of the six degree of freedom response motions of a floating wind turbine in ocean waves. Coupling these models with near-field real time irregular sea surface (wve) measurement/sensings and prediction models, we used the digital twin to predict how the floating turbine will respond to the incoming waves. Applying this information to a motion control system of the float, one can limit and control float motions to prevent undesirable loading events/large angular motions, thus increasing system life and ultimately contributing to optimizing energy production. Due to the computational intensity of operating a digital twin in real time, we investigated the use of artificial intelligence techniques to speed-up the processes of the digital twin, as well as the wave reconstruction/prediction models. Model tank testing at the University of Rhode Island and University of Maine both validated and demonstrated the developed techniques on simple float geometries and a scale model of the NREL 15 MW reference turbine.

17 WIND ENERGY↗

Parallelization techniques for quantum simulation of fermionic systems

Mapping fermionic operators to qubit operators is an essential step for simulating fermionic systems on a quantum computer. We investigate how the choice of such a mapping interacts with the underlying qubit connectivity of the quantum processor to enable (or impede) parallelization of the resulting Hamiltonian-simulation algorithm. It is shown that this problem can be mapped to a path coloring problem on a graph constructed from the particular choice of encoding fermions onto qubits and the fermionic interactions onto paths. The basic version of this problem is called the weak coloring problem. Taking into account the fine-grained details of the mapping yields what is called the strong coloring problem, which leads to improved parallelization performance. A variety of illustrative analytical and numerical examples are presented to demonstrate the amount of improvement for both weak and strong coloring-based parallelizations. Our results are particularly important for implementation on near-term quantum processors where minimizing circuit depth is necessary for algorithmic feasibility.

97 MATHEMATICS AND COMPUTING↗

Update on Parallel Process Execution in the Next Generation System Analysis Model

As of the end of 2021, 88,880 metric tons of heavy metal (MTHM) (44,741 MTHM in dry storage; 44,139 MTHM in wet storage) of spent nuclear fuel (SNF) were stored at various reactor sites across the United States [1]. The Office of Storage and Transportation in the Department of Energy is planning for the transportation, storage, and eventual disposal of SNF and high-level radioactive waste (HLW). To aid in this effort and inform decision-makers about the backend of the spent fuel cycle, systems analysis tools capable of analyzing the various options with respect to SNF and HLW management are being used as well as continuously improved to meet the evolving needs of the program. System analysts typically use these tools to vary underlying assumptions (shipping rates, allocation priority, available facilities, start dates, etc.) and study the implications of these changes on site clearance schedules, campaign costs, transportation infrastructure acquisition, etc.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

A two-level GPU-accelerated incomplete LU preconditioner for general sparse linear systems

This paper presents a parallel preconditioning approach based on incomplete LU (ILU) factorizations in the framework of Domain Decomposition (DD) for general sparse linear systems. We focus on distributed memory parallel architectures, specifically, those that are equipped with graphic processing units (GPUs). In addition to block-Jacobi, we present general purpose two-level ILU Schur complement-based approaches, where different strategies are presented to solve the coarse-level reduced system. These strategies are combined with modified ILU methods in the construction of the coarse-level operator, in order to effectively remove smooth errors by targeting an algebraically smooth vector. We leverage available GPU-based sparse matrix kernels to accelerate the setup and the solve phases of the proposed ILU preconditioner. We evaluate the efficiency of the proposed methods as a smoother for algebraic multigrid (AMG) and as a preconditioner for Krylov subspace methods on challenging anisotropic diffusion problems and a collection of general sparse matrices.

97 MATHEMATICS AND COMPUTING↗

Feedback lock-in: A versatile multi-terminal measurement system for electrical transport devices

Here, we present the design and implementation of a measurement system that enables parallel drive and detection of small currents and voltages at numerous electrical contacts to a multi-terminal electrical device. This system, which we term a feedback lock-in, combines digital control-loop feedback with software-defined lock-in measurements to dynamically source currents and measure small, pre-amplified potentials. The effective input impedance of each current/voltage probe can be set via software, permitting any given contact to behave as an open-circuit voltage lead or as a virtually grounded current source/sink. This enables programmatic switching of measurement configurations and permits measurement of currents at multiple drain contacts without the use of current preamplifiers. Our 32-channel implementation relies on commercially available digital input/output boards, home-built voltage preamplifiers, and custom open-source software. With our feedback lock-in, we demonstrate differential measurement sensitivity comparable to a widely used commercially available lock-in amplifier and perform efficient multi-terminal electrical transport measurements on twisted bilayer graphene and SrTiO 3 quantum point contacts. The feedback lock-in also enables a new style of measurement using multiple current probes, which we demonstrate on a ballistic graphene device.

47 OTHER INSTRUMENTATION↗

Massively Parallel Capability in Sierra/SD for Simulation Vibration with Piezoelectrics

Sierra/SD is an engineering structural dynamics code that provides Sandia and other customers a tool to model structural and acoustic physics on large complex physical systems using massively parallel processing. This report provides a detailed overview on Sierra/SD’s most recent physics package: coupled electro-mechanical physics. This capability uses the finite element method to model coupled electro-mechanical physics exhibited by piezoelectric materials. This report provides an applications overview, theory overview, and verification examples demonstrating the electro-mechanical physics modeling capabilities of Sierra/SD.

97 MATHEMATICS AND COMPUTING↗

TEAM Project Review, Year 2

This report summarizes our research activities within the TEAM project between December 2020 and December 2021, funded by the ASCR Advanced Research in Quantum Computing program. During the reporting period the LLNL-MSU team has made progress on several fronts. An overarching goal of the team is to provide a comprehensive suite of software tools that can be used for the Characterize-Optimize-Compute loop needed to implement and execute algorithms on quantum devices. We are concurrently developing lightweight solvers that can be used on desktop computers to find optimal control pulses and to characterize small quantum systems (consisting of a few transmons and cavities). However, desktop computers are insufficient for simulating and characterizing larger quantum systems. We have therefore also developed parallel, distributed memory, simulators and optimization solvers, both for open and closed quantum systems. These parallel solvers have, for example, been used to study quantum optimal control for pure-state preparation, utilizing 1000’s of cores on a modern high-performance computing (HPC) platform.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Development, construction and tests of the Mu2e electromagnetic calorimeter mechanical structures

The “muon-to-electron conversion” (Mu2e) experiment at Fermilab will search for the charged lepton flavour violating neutrino-less coherent conversion of a muon into an electron in the field of an aluminum nucleus. The observation of this process would be the unambiguous evidence of the existence of physics beyond the standard model. Mu2e detectors comprise a straw-tracker, an electromagnetic calorimeter and an external veto for cosmic rays. In particular, the calorimeter provides excellent electron identification, a fast calorimetric online trigger, and complementary information to aid pattern recognition and track reconstruction. The detector has been designed as a state-of-the-art crystal calorimeter and employs 1348 pure Cesium Iodide (CsI) crystals readout by UV-extended silicon photosensors and fast front-end and digitization electronics. A design consisting of two identical annular matrices (named “disks”) positioned at the relative distance of 70 cm downstream the aluminum target along the muon beamline satisfies the Mu2e physics requirements. The hostile Mu2e operational conditions, in terms of radiation levels (total expected ionizing dose of 12 krad and a neutron fluence of 5 × 10$^{10}$ n/cm$^{2}$ @ 1 MeV$_{eq}$ (Si)/y), magnetic field intensity (1 T) and vacuum level (10$^{-4}$ Torr) have posed tight constraints on scintillating materials, sensors, electronics and on the design of the detector mechanical structures and material choice. The support structure of each 674 crystal matrix is composed of an aluminum hollow ring and parts made of open-cell vacuum-compatible carbon fiber. The photosensors and front-end electronics for the readout of each crystal are inserted in a machined copper holder and make a unique mechanical unit. The resulting 674 mechanical units are supported by a machined plate of vacuum-compatible plastic material. The plate also integrates the cooling system made of a network of copper lines flowing a low temperature radiation-hard fluid and placed in thermal contact with the copper holders to constitute a low resistance thermal bridge. The data acquisition electronics are hosted in aluminum custom crates positioned on the external lateral surface of the disks. The crates also integrate the electronics cooling system as lines running in parallel to the front-end system. In this paper we report on the calorimeter mechanical structure design, the mechanical and thermal simulations that have determined the design technological choices, and the status of component production, quality assurance tests and plans for assembly at Fermilab.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

DIMPLES: Distributed Influence Maximization for Pandemic pLanning on Exascale Systems

We study exascale parallel algorithms for the selection of intervention or monitoring strategies in massive realistic socio-technical networks through scalable Influence Maximization (InfMax) algorithms. We employ novel techniques to enable efficient scaling on up to 8k nodes of OLCF Frontier, with 65k AMD GPUs and 458k AMD CPU cores. Current state-of-the-art InfMax tools are limited to networks with only a few million actors (vertices) and a few hundred million interactions (edges). By overcoming these limitations, we show that our approach is capable of processing a realistic social contact network of the United States with 285 million nodes and about 8 billion edges. This two orders-of-magnitude improvement over the previous state-of-the-art is obtained by leveraging algorithmic advancements for the InfMax problem and designing several problem-specific approaches to overlap communication with computation, improve GPU efficiency, and lower the application’s memory requirements. We evaluate strong scaling for computing 10k most influential seeds using up to 8k nodes of an exascale system, and weak scaling from 128 to 8k system nodes for seed sets ranging from 625 to 40k seeds. We achieve the fastest-known runtime of 25 minutes while performing 48 million diffusion simulations totaling 2.31 petabytes to identify 40k influential seeds using 8k nodes, and take 5.75 minutes to identify 10k seeds while using 4k nodes.

Minutoli, Marco [Pacific Northwest National Labora↗

Testing and Analysis of Grid Forming Inverter Control for Achieving Resilient and Economic Operation of an Islanded Microgrid

This investigation examines the feasibility of operating a battery energy storage system (BESS) in parallel with synchronous generation by using grid forming (GFM) control in order to achieve frequency control objectives while mitigating increases to operating costs in the context of an islanded microgrid. The BESS GFM control system, which is based on conventional droop techniques, is modeled along with the overall microgrid using the Real Time Digital Simulator (RTDS) to allow for integration of genset controller hardware. A series of simulations are performed to test the voltage and frequency regulation capability of the BESS control system when the primary frequency regulating genset is tripped offline. The results of the simulations suggest that the GFM control scheme will successfully maintain frequency and voltage stability, which will enable operation without a back-up genset while not compromising the microgrid resiliency to contingencies.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Ray-tracing analysis for cross-polarization scattering diagnostic on MAST-upgrade spherical tokamak

A combined Doppler backscattering/cross-polarization scattering (DBS/CPS) system is being deployed on MAST-U for simultaneous measurements of local density turbulence, turbulence flows, and magnetic turbulence. In this design, CPS shares the probing beam with the DBS and uses a separate parallel-viewing receiver system. In this study, we utilize a modified GENRAY 3D ray-tracing code to simulate the propagation of the probing and scattered beams. The contributions of different scattering locations along the entire beam trajectories are considered, and the corresponding local B̃ wavenumbers are estimated using the wavevector matching criterion. The wavenumber ranges of the local B̃ that are detectable to the CPS system are explored for simulated L- and H-mode plasmas.

Hong, R. (ORCID:0000000347508015)↗

A Robust Parallel Distributed State Estimation for Large Scale Distribution Systems

The growing need and interest in real-time monitoring of large distribution networks motivated by the rapid population of renewable sources, EVs and etc. demand a computationally efficient state estimation framework. Furthermore, this paper presents an improved computational framework for implementing a robust state estimator using a multi-core processor. The main contribution of the paper is the proposed computational framework along with two partitioning strategies which enable fast and robust state estimation for large scale radial and/or meshed distribution systems. Formulation of the proposed method and its implementation are described in detail. Performance of the estimator is tested by simulations first using a small 84-bus radial distribution system. Then the method’s scalability is demonstrated by simulations on two very large scale distribution networks one configured radially and the other meshed each containing over 12,500 buses.

42 ENGINEERING↗

Game Theoretic Orchestration for Cooperation among Power Distribution System Applications

The evolving transformation with the proliferation of distributed energy resources and advanced metering, necessitates advanced distribution systems to integrate and orchestrate a large number of grid-edge devices while also serving multiple system-level objectives such as resilience, decarbonization, equity and other system mandates. The parallel deployment and control of resources towards achieving diverse objectives may lead to conflicts between applications that want to control overlapping sets of device setpoints, potentially leading to oscillatory behavior and suboptimal performance. This work aims at leveraging game theoretic framework to drive cooperative behavior among competitive applications. The work proposes a weighted-consensus based game design to facilitate conflict resolution through consensus-building iterations for modular platform. Simulation-based evaluation on a sample test system demonstrates the performance the proposed deconfliction strategy in resolving operational conflicts and achieving close-to-optimal trade off among the applications. Results also compare the proposed strategy with a distribution optimization approach and illustrate it effectiveness in diverse apps regardless of their design while also incentivizing apps with flexible design.

Advanced distribution operations, cooperation, app↗

Understanding Lustre Internals. Second Edition

The Lustre file system has become a preferred storage resource for systems on the Top500 list, and it is often the file system of choice for small- to medium-sized HPC systems that require parallel shared access to data. Several resources exist to help users deploy and configure Lustre, but the same cannot be said for resources that explain the inner workings of the Lustre source code. A previous ORNL technical report entitled "Understanding Lustre Filesystem Internals" (ORNL/TM-2009/117) provided an excellent summary of Lustre subsystem operations. However, that report is over a decade old and is based on Lustre version 1.6. Since that report was published, Lustre has evolved significantly. Several subsystems underwent significant code changes and many new features have been added to the file system, bringing the current Lustre version up to 2.15.This report aims to document and explain the internal workings of the latest version of the Lustre file system. It will provide more complete and up-to-date information than the previous technical report and should serve as a foundational document for anyone interested in Lustre software development. Key data structures will be described along with the APIs used for interaction among the various Lustre subsystems. Although the Lustre software is constantly being developed, the details in this document should remain relevant for the forseeable future.

97 MATHEMATICS AND COMPUTING↗

New Horizons for High-Performance Computing

Here we provide an overview of the past, present, and a diverse collection of future computer architecture alternatives for HPC. The end of Moore’s Law influenced the current HPC architecture focus on accelerated compute nodes composed of CPU and GPU computing components integrated into massively parallel processor architecture systems. There are many alternatives for future HPC directions, with different technologies, computing ecosystems, opportunities for lead user application-driven customization, and the role of open innovation business models. This paper provides an overview of these different new horizons for HPC, an organizing principle to focus future computing research, different public-private partnership models, and the critical role of workforce development.

97 MATHEMATICS AND COMPUTING↗

System and method of storing and analyzing information

A system and method of storing and analyzing information is disclosed. The system includes a compiler layer to convert user queries to data parallel executable code. The system further includes a library of multithreaded algorithms, processes, and data structures. The system also includes a multithreaded runtime library for implementing compiled code at runtime. The executable code is dynamically loaded on computing elements and contains calls to the library of multithreaded algorithms, processes, and data structures and the multithreaded runtime library.

Feo, John T.↗