Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “distributed parallelization”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 415 records · Page 23

Bifurcation-like transition of divertor conditions induced by X-point radiation in KSTAR L-mode plasmas *

Abstract Density ramps with ion grad B drift directed into lower single null KSTAR L-mode plasmas are associated with a simultaneous and abrupt reduction of the divertor particle flux on both low- and high-field-side targets when the mid-plane line averaged electron density reaches a given level. Target embedded Langmuir probe signals show a clear ‘cliff edge’ behavior similar to that observed in the divertor target electron temperature in DIII-D H-mode plasmas (Eldon et al 2017 Nucl. Fusion 57 066039; McLean et al 2015 J. Nucl. Mater. 463 533–6). The collapse of the particle flux is observed along the whole divertor target area (from private flux region to the far scrape-off layer (SOL)). The critical upstream density of this target flux cliff is invariant under fuel gas throughput modulation. The transition along the cliff occurs in tens of milliseconds. With the cliff, carbon impurities and deuterium neutrals transported through the X-point to the core produce a strong radiation spot near the X-point, seen on bolometric signals, and increase the upstream density. The experimental observations are consistent with time-dependent SOLPS-ITER simulations, which also demonstrate an abrupt transition of the target flux and upstream density with the increase in X-point radiation. The timescale of the cliff predicted by SOLPS-ITER is consistent with the experiment, although, it is influenced by gas throughput or time-dependent numerical methods. In the L-mode phase space of separatrix electron density and temperature, branches are divided based on target temperature, because the latter is strongly coupled to the radiation front and ionization front due to the monotonic characteristic of the parallel electron temperature distribution. Since the H-mode condition operates at a much higher upstream density and electron temperature in phase space, dissipation from sputtered carbon alone leads to the density limit before reaching the X-point radiation condition. This is therefore consistent with the fact that cliffs have never been observed in H-mode KSTAR experiments.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

A semi-automated algorithm for designing stellarator divertor and limiter plates and application to HSX

We present a semi-automated algorithm for designing three-dimensional divertor or limiter plates targeting low heat loads. The algorithm designs the plates in two stages: firstly, the parallel heat flux distribution is caught on vertically-inclined plates at one or several toroidal locations. Secondly, the power per unit area is reduced by stretching, tilting and bending the plates toroidally. Heat transport is modelled using the EMC3-Lite code, which uses an anisotropic diffusion model. We apply this scheme to HSX, a medium-sized stellarator located at the University of Wisconsin–Madison. Starting from the current machine with an extended vessel wall, we construct plates which are able to effectively catch and spread the heat for three different magnetic configurations. The scheme has a computational cost in the order of tens of CPU-minutes, making it a powerful tool for semi-automated plasma-facing component design in three-dimensional environments.

anisotropic diffusion↗

Grid-Forming and Grid-Following Inverter Comparison of Droop Response

With the increase in penetration of inverter-based resources (IBRs) in the electrical power system, the ability of these devices to provide grid support to the system has become a necessity. With standards previously developed for the interconnection requirements of grid-following inverters (GFLI) (most commonly photovoltaic inverters), it has been well documented how these inverters “should” respond to changes in voltage and frequency. However, with other IBRs such as grid-forming inverters (GFMIs) (used for energy storage systems, standalone systems, and as uninterruptable power supplies) these requirements are either: not yet documented, or require a more in deep analysis. With the increased interest in microgrids, GFMIs that can be paralleled onto a distribution system have become desired. With the proper control schemes, a GFMI can help maintain grid stability through fast response compared to rotating machines. This paper will present an experimental comparison of commercially available GFMI and GFLI ' responses to voltage and frequency deviation, as well as the GFMI operating as a standalone system and subjected to various changes in loads.

Grid Support, Inverter, Droop Control, Volt-VAR, F↗

Optimizing the Weather Research and Forecasting Model with OpenMP Offload and Codee

Currently, the Weather Research and Forecasting model (WRF) utilizes shared memory (OpenMP) and distributed memory (MPI) parallelisms. To take advantage of GPU resources on the Perlmutter supercomputer at NERSC, we port parts of the computationally expensive routine Fast Spectral Bin Microphysics (FSBM) to NVIDIA GPUs using OpenMP device offloading directives. To facilitate this process, we explore a workflow for optimization which uses both runtime profilers and a static code inspection tool Codee to refactor the subroutine. We observe an 2.24x overall speedup for the CONUS-12km storm test case.

Wichitrnithed, Chayanon (Namo) [Odin Institute]↗

Quandary

Quandary numerically simulates and optimizes the time-evolution of open quantum systems. The underlying dynamics are modelled by Lindblad's master equation, a linear ordinary differential equation (ODE) describing quantum systems interacting with the environment. Quandary solves this ODE numerically by applying a time-stepping integration scheme, and utilizes a gradient-based optimization approach to determine optimal control pulses that drive the quantum system to a desired target state. Two optimization objectives are considered: (a) Unitary gate optimization that finds controls to realize a unitary gate transformation, and (b) optimal reset that aims to drive the quantum system to the ground states. Gradient-based optimization schemes utilizing Petsc's Tao optimization package are applied to generate control pulses that minimize the respective measure. To evaluate the gradient of the objective function, the discrete adjoint method is used while leveraging techniques from Algorithmic Differentiation to produce exact and consistent gradients. To mitigate excessive execution run times, the software can be build together with the XBraid software library which provides a parallelization strategy to distribute the time-evolution of the underlying dynamics onto multiple processor.

Petersson, NilsA.↗

Decomposition and Algorithmic Approaches for Solving Large-Scale Process Family Design Problems

Our most recent work expands the water desalination case study from 76 variants to 10,897 variants using the equation-oriented model built in Pyomo as part of the PARETO project. Using the discretization formulation presented in Stinchfield (2024a), rather than solving for all 10,897 variants simultaneously, we decompose the formulation into subproblems containing subsets of variants from the process family. We solve the overall problem with Progressive Hedging (PH) deployed in parallel on a distributed HPC cluster using the open-source Python package mpi-sppy (Knueven et al., 2023). This approach allowed us to solve this process family design problem to ~1.5% relative optimality gap in about 5 hours; in comparison, Gurobi reached ~50% relative optimality gap in about 6 hours (Stinchfield et al., 2024b). However, this approach still requires discretization of the common unit module design ranges; additionally, PH acts as a heuristic for MILP’s with gap-closing capabilities. Ideally, we would not have to use ML surrogates or discretization to solve this problem, instead solving the process family design problem with the equation-oriented model directly to achieve the most accurate results. However, recall that we did not consider solving the MINLP directly due to complexity and size. In this work, we aim to decompose and solve this large-scale MINLP using a Structured Nonlinear Global Optimization algorithm presented by Cao and Zavala (2019).

Stinchfield, Georgia↗

On bottom density currents on the continental shelves

The turbulent characteristics of bottom density currents on the continental shelves and their influence on the vertical profiles of current velocities are studied by considering plane parallel flows of a liquid with one density in a motionless liquid and with lighter density along an inclined plane. The motion of the liquid is a result of gravitational force directed along the parallel plane. Vertical distribution of turbulent stress is determined from a known average velocity profile and is used to obtain the vertical profile of the average current velocity.

Anuchin, V. N.↗

RAMP: A fault tolerant distributed microcomputer structure for aircraft navigation and control

RAMP consists of distributed sets of parallel computers partioned on the basis of software and packaging constraints. To minimize hardware and software complexity, the processors operate asynchronously. It was shown that through the design of asymptotically stable control laws, data errors due to the asynchronism were minimized. It was further shown that by designing control laws with this property and making minor hardware modifications to the RAMP modules, the system became inherently tolerant to intermittent faults. A laboratory version of RAMP was constructed and is described in the paper along with the experimental results.

Dunn, W. R.↗

Corrected formula for the polarization of second harmonic plasma emission

Corrections for the theory of polarization of second harmonic plasma emission are proposed. The nontransversality of the magnetoionic waves was not taken into account correctly and is here corrected. The corrected and uncorrected results are compared for two simple cases of parallel and isotropic distributions of Langmuir waves. It is found that whereas with the uncorrected formula plausible values of the coronal magnetic fields were obtained from the observed polarization of the second harmonic, the present results imply fields which are stronger by a factor of three to four.

Melrose, D. B.↗

A simplified approach to axisymmetric dual-reflector antenna design

A procedure is described for designing dual reflector antennas. The analysis is developed by taking each reflector to be the envelope of its tangent planes. Rather than specifying the phase distribution in the emitted beam, the slopes of the emitted rays were specified. Thus, both the output wave shape and angular distribution of intensity can be specified. Computed examples include variations from both Cassegrain and Gregorian systems, permitting deviation from uniform source distributions and from parallel beam property of conventional systems.

Barger, Raymond L.↗

Decentralized Adaptive Control For Robots

Precise knowledge of dynamics not required. Proposed scheme for control of multijointed robotic manipulator calls for independent control subsystem for each joint, consisting of proportional/integral/derivative feedback controller and position/velocity/acceleration feedforward controller, both with adjustable gains. Independent joint controller compensates for unpredictable effects, gravitation, and dynamic coupling between motions of joints, while forcing joints to track reference trajectories. Scheme amenable to parallel processing in distributed computing system wherein each joint controlled by relatively simple algorithm on dedicated microprocessor.

Seraji, Homayoun↗

Extremely high data-rate, reliable network systems research

Significant progress was made over the year in the four focus areas of this research group: gigabit protocols, extensions of metropolitan protocols, parallel protocols, and distributed simulations. Two activities, a network management tool and the Carrier Sensed Multiple Access Collision Detection (CSMA/CD) protocol, have developed to the point that a patent is being applied for in the next year; a tool set for distributed simulation using the language SIMSCRIPT also has commercial potential and is to be further refined. The year's results for each of these areas are summarized and next year's activities are described.

Foudriat, E. C.↗

A discrete decentralized variable structure robotic controller

A decentralized trajectory controller for robotic manipulators is designed and tested using a multiprocessor architecture and a PUMA 560 robot arm. The controller is made up of a nominal model-based component and a correction component based on a variable structure suction control approach. The second control component is designed using bounds on the difference between the used and actual values of the model parameters. Since the continuous manipulator system is digitally controlled along a trajectory, a discretized equivalent model of the manipulator is used to derive the controller. The motivation for decentralized control is that the derived algorithms can be executed in parallel using a distributed, relatively inexpensive, architecture where each joint is assigned a microprocessor. Nonlinear interaction and coupling between joints is treated as a disturbance torque that is estimated and compensated for.

Tumeh, Zuheir S.↗

Hyperswitch communication network

The Hyperswitch Communication Network (HCN) is a large scale parallel computer prototype being developed at JPL. Commercial versions of the HCN computer are planned. The HCN computer being designed is a message passing multiple instruction multiple data (MIMD) computer, and offers many advantages in price-performance ratio, reliability and availability, and manufacturing over traditional uniprocessors and bus based multiprocessors. The design of the HCN operating system is a uniquely flexible environment that combines both parallel processing and distributed processing. This programming paradigm can achieve a balance among the following competing factors: performance in processing and communications, user friendliness, and fault tolerance. The prototype is being designed to accommodate a maximum of 64 state of the art microprocessors. The HCN is classified as a distributed supercomputer. The HCN system is described, and the performance/cost analysis and other competing factors within the system design are reviewed.

Peterson, J.↗

Programming in a proposed 9X distributed Ada

The proposed Ada 9X constructs for distribution was studied. The goal was to select suitable test cases to help in the evaluation of the proposed constructs. The examples were to be considered according to the following requirements: real time operation; fault tolerance at several different levels; demonstration of both distributed and massively parallel operation; reflection of realistic NASA programs; illustration of the issues of configuration, compilation, linking, and loading; indications of the consequences of using the proposed revisions for large scale programs; and coverage of the spectrum of communication patterns such as predictable, bursty, small and large messages. The first month was spent identifying possible examples and judging their suitability for the project.

Waldrop, Raymond S.↗

A prototype heat pipe heat exchanger for the capillary pumped loop flight experiment

A Capillary Pumped Two-Phase Heat Transport Loop (CAPL) Flight Experiment, currently planned for 1993, will provide microgravity verification of the prototype capillary pumped loop (CPL) thermal control system for EOS. CAPL employs a heat pipe heat exchanger (HPHX) to couple the condenser section of the CPL to the radiator assembly. A prototype HPHX consisting of a heat exchanger (HX), a header heat pipe (HHP), a spreader heat pipe (SHP), and a flow regulator has been designed and tested. The HX transmits heat from the CPL condenser to the HHP, while the HHP and SHP transport heat to the radiator assembly. The flow regulator controls flow distribution among multiple parallel HPHX's. Test results indicated that the prototype HPHX could transport up to 800 watts with an overall heat transfer coefficient of more than 6000 watts/sq m-deg C. Flow regulation among parallel HPHX's was also demonstrated.

Ku, Jentung↗

Quasi-optical Josephson-junction oscillator arrays

Josephson junctions are natural voltage-controlled oscillators capable of generating submillimeter-wavelength radiation, but a single junction usually can produce only 100 nW of power and often has a broad spectral linewidth. The authors are investigating 2D quasi-optical power combining arrays of 103 and 104 NbN/MgO/NbN and Nb/Al-AlO(x)/Nb junctions to overcome these limitations. The junctions are dc-biased in parallel and are distributed along interdigitated lines. The arrays couple to a resonant mode of a Fabry-Perot cavity to achieve mutual phase-locking. The array configuration has a relatively low impedance, which should allow the capacitance of the junctions to be tuned out at the oscillation frequency.

Stern, J. A.↗

Computational Methods for HSCT-Inlet Controls/CFD Interdisciplinary Research

A program aimed at facilitating the use of computational fluid dynamics (CFD) simulations by the controls discipline is presented. The objective is to reduce the development time and cost for propulsion system controls by using CFD simulations to obtain high-fidelity system models for control design and as numerical test beds for control system testing and validation. An interdisciplinary team has been formed to develop analytical and computational tools in three discipline areas: controls, CFD, and computational technology. The controls effort has focused on specifying requirements for an interface between the controls specialist and CFD simulations and a new method for extracting linear, reduced-order control models from CFD simulations. Existing CFD codes are being modified to permit time accurate execution and provide realistic boundary conditions for controls studies. Parallel processing and distributed computing techniques, along with existing system integration software, are being used to reduce CFD execution times and to support the development of an integrated analysis/design system. This paper describes: the initial application for the technology being developed, the high speed civil transport (HSCT) inlet control problem; activities being pursued in each discipline area; and a prototype analysis/design system in place for interactive operation and visualization of a time-accurate HSCT-inlet simulation.

Cole, Gary L.↗