Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “asynchronous”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 127 records · Page 7

Stability and control of power systems with high penetrations of inverter-based resources: An accessible review of current knowledge and open questions

As power system renewable energy penetrations increase, the ways in which key renewable technologies such as wind and solar photovoltaics (PV) differ from thermal generators become more apparent. Many studies have examined the variability and uncertainty of such generators and described how generation and load can be balanced for a wide variety of annual energy penetrations, at timescales from seconds to years. Another important characteristic of these resources is asynchronicity, the result of using inverters to interface the prime energy source with the power system as opposed to synchronous generators. Unlike synchronous generators, whose frequency of alternating current (AC) injection is physically coupled to the rotation of the machine itself, inverter based asynchronous generators do not share the same physical coupling with the generated frequency. These subtle differences impact the operations of power systems developed around the characteristics of synchronous generators. In this paper we review current knowledge and open research questions concerning the interplay between asynchronous inverter-based resources (IBRs) and cycle- to second-scale power system dynamics, with a focus on how stability and control may be impacted or need to be achieved differently when there are high instantaneous penetrations of IBRs across an interconnection. This work does not seek to provide a comprehensive review of the latest developments, but is instead intended to be accessible to any reader with an engineering background and an interest in power systems and renewable energy. As such, the paper includes basic material on power electronics, control schemes for IBRs, and power system stability; and uses this background material to describe potential impacts of IBRs on power system stability, operational challenges associated with large amounts of distributed IBR generation, and modern power system simulation trends driven by IBR characteristics.

14 SOLAR ENERGY↗

Trigger-based Incremental Data Processing with Unified Sync and Async Model

In recent years, more and more applications in the cloud have needs to process large-scale on-line datasets, which evolve over time as new entries are added and existing entries are modified. Several programming frameworks, such as Percolator and Oolong, are proposed for such incremental data processing and can achieve efficient processing with an event-driven abstraction. However, these frameworks are inherently asynchronous, leaving the heavy burden of managing synchronization to applications' developers, which further significantly restricts their usabilities. In this study, we propose a trigger-based incremental computing framework in the cloud, called Domino, with both synchronous and asynchronous mechanisms to coordinate parallel triggers. With this new framework, both synchronous and asynchronous applications can be seamlessly developed. Use cases and extensive evaluation results confirm that it can deliver sufficient performance, and also is easy to use for incremental applications in large-scale distributed computing.

97 MATHEMATICS AND COMPUTING↗

Online Distribution System State Estimation via Stochastic Gradient Algorithm

Distribution network operation is becoming more challenging because of the growing integration of intermittent and volatile distributed energy resources (DERs). This motivates the development of new distribution system state estimation (DSSE) paradigms that can operate at fast timescale based on real-time data stream of asynchronous measurements enabled by modern information and communications technology. To solve the real-time DSSE with asynchronous measurements effectively and accurately, this paper formulates a weighted least squares DSSE problem and proposes an online stochastic gradient algorithm to solve it. The performance of the proposed scheme is analytically guaranteed and is numerically corroborated with realistic data on IEEE 123-bus feeder.

distribution system state estimation↗

HDF5 in the exascale era: Delivering efficient and scalable parallel I/O for exascale applications

Accurately modeling real-world systems requires scientific applications at exascale to generate massive amounts of data and manage data storage efficiently. However, parallel input and output (I/O) faces challenges due to new application workflows and the state-of-the-art memory, interconnect, and storage architectures considered in exascale designs. The storage hierarchy has expanded with node-local persistent memory, solid-state storage, and traditional disk and tape-based storage, thus requiring efficiency at each layer and much more efficient data movement among these layers. This paper discusses how the ExaHDF5 project improved the I/O performance and data management for exascale architectures by enhancing HDF5, a widely used parallel I/O library. The team developed an Asynchronous I/O Virtual Object Layer (VOL) connector that allowed overlapping I/O with computation. They also created a Cache VOL to complement asynchronous I/O by incorporating fast storage layers, such as burst buffer and node-local storage, into the parallel I/O workflow through caching and staging data. Additionally, the team enabled data aggregation and I/O at the node level by using a Subfiling Virtual File Driver (VFD). To demonstrate superior I/O performance with HDF5 at exascale, the ExaHDF5 team collaborated with several exascale applications. In this paper, we show I/O performance improvements for three applications: Cabana (a particle-based simulation library), EQSIM (a regional earthquake simulation software), and E3SM (a climate system modeling library).

Asynchronous I/Ol↗

Benefits of high-voltage SiC-based power electronics in medium-voltage power-distribution grids

Medium-voltage (MV) power electronics equipment has been increasingly applied in distribution grids, and high-voltage (HV) silicon carbide (SiC) power semiconductors have attracted considerable attention in recent years. This paper first overviews the development and status of HV SiC power semiconductors. Then, MV power-converter applications in distribution grids are summarized and the benefits of HV SiC in these applications are presented. Microgrids, including conventional and asynchronous microgrids, that can fully demonstrate the benefits of HV SiC power semiconductors are selected to investigate the benefits of HV SiC in detail, including converter-level benefits and system-level benefits. Finally, an asynchronous microgrid power-conditioning system (PCS) prototype using a 10 kV SiC MOSFET is presented.

24 POWER TRANSMISSION AND DISTRIBUTION↗

SVM-Based Synchronized Fault Detection for 100% Renewable Microgrids

Traditional protection schemes face significant challenges when applied to microgrids with high penetrations of renewables with inverter-based resources (IBRs). The proliferation of advanced sensing and communication technologies has generated copious data, offering an opportunity to overcome these limitations using data-driven machine learning approaches. This work proposes a novel approach based on a support vector machine (SVM) for detecting faults within a 100% renewable microgrid. The approach encompasses a systematic offline training stage for the development of a linear SVM-based fault detection algorithm. This process covers offline data collection from the microgrid under study, the extraction of features such as positive- and negative-sequence components and the total harmonic distortion of the voltage and current measurements of the relays, and the design of the linear SVM-based classifier. During the online implementation, however, different classifiers can exhibit asynchronicity in detecting the fault inception at different subcycle-to-cycle period-level delays. To circumvent this asynchronicity issue, a separate algorithm is developed for each relay to estimate the fault inception time as close to the real fault time. The performance of the proposed SVM-based synchronized fault detection method is evaluated using online time-domain simulation studies on a microgrid test system. The results corroborate the reliability of the fault detection scheme when tested under various fault cases (fault types, locations, and impedances) and non-fault cases during both grid-tied and islanded operation modes.

100% microgrid↗

Peak Reduction Using Mode Adjustment of Heat Pump Water Heaters in a Residential Neighborhood

Building electrification is putting pressure on distribution grid worldwide. Peak reduction is an important concern that can help reduce the growing stress and allow to defer investments in new capacity. Water heaters represent a convenient way of reducing peak because they are less dependent on weather, and their storage volume allows for asynchronous water heating and hot water use. Previous empirical studies investigated the ability of water heaters to reduce peak through the adjustment of the temperature setpoint. However, not all equipment vendors offer this option. This study aims at understanding the feasibility of peak reduction with an alternative configuration available in the market - by adjusting the device mode rather than temperature setpoint. The peak reduction methodology is tested in an occupied 46-townhome neighborhood located in Atlanta, GA. We find that peak shifting is possible with the adjustable mode approach, with the change in the peak load by 30-60%.

demand response↗

Relative reproductive phenology and synchrony affect neonate survival in a nonprecocial ungulate

Abstract Degree of reproductive synchronization in prey is hypothesized as a predator defense strategy reducing prey risk via predator satiation or predator avoidance. Species with precocial young, especially those exposed to specialist predators, should be highly synchronous to satiate predators (predator satiation hypothesis), while prey with nonprecocial (i.e. altricial) young, especially those exposed to generalist predators, should become relatively asynchronous to avoid predator detection (predator avoidance hypothesis). The white‐tailed deer Odocoileus virginianus in North America is an example of a nonprecocial ungulate that uses the hider strategy early in life; its primary predator (coyote; Canis latrans ) is a generalist, making white‐tailed deer a good model species to test the predator avoidance hypothesis. We used birth dates and known fates of white‐tailed deer neonates ( n = 1,032) across nine study sites varying in relative synchrony and predator assemblages to test the predator avoidance hypothesis. We predicted that relative birthing asynchrony of the population would increase relative survival at the population level; therefore, at the individual scale, neonate birth date nearer to mean birthing date in a respective population would not influence individual survival. Coyotes were responsible for the majority of predation events, and survival of those neonates increased the closer the individual was born to peak birthing season in each respective population. Also, at the population level, reproductive asynchronization negatively affected survival. Contrary to the predator avoidance hypothesis, our data indicate patterns in neonate survival for white‐tailed deer better support the predator satiation hypothesis at the individual and population level. Additionally, coyotes may present a selective force great enough to shift reproductive synchrony such that predator satiation may become a feasible defense strategy for neonates at local spatial scales. Our results indicate that synchronizing reproduction may still be the most effective strategy to reduce individual predation risk from generalist predators, particularly when the window of heightened resource availability to the prey is narrow. A free Plain Language Summary can be found within the Supporting Information of this article.

Michel, Eric S.↗

Electrode strain dynamics in layered intercalation battery cathodes

Rechargeable batteries using electrodes based on intercalation chemistry exhibit notable cyclability, yet their performance still suffers from chemomechanical degradation. In this study, by combining a suite of operando microscopy methods, we explored electrode strain evolution and observed intricate particle cluster rearrangement under electrochemical stimuli. We show that early-stage strain accumulation in intercalation cathodes occurs during the period of interparticle charge transfer and redox reactions stemming from asynchronous coupling and decoupling between chemical (de)intercalation and physical grain motion. This interplay drives heterogeneous redox activity, localized charge equilibration, and multiscale strain cascades that propagate through an asynchronous network of chemical-mechanical interactions. Together, these findings reveal how collective particle dynamics and hierarchical strain transmission dictate electrode deformation and degradation in intercalation cathodes.

25 ENERGY STORAGE↗

The Minos Computing Library: Efficient Parallel Programming for Extremely Heterogeneous Systems

Hardware specialization has become the silver bullet to achieve efficient high performance, from Systems-on-Chip systems, where hardware specialization can be ``extreme'', to large-scale HPC systems. As the complexity of the systems increases, so does the complexity of programming such architectures in a portable way. This work introduces the Minos Computing Library (MCL), as system software, programming model, and programming model runtime that facilitate programming extremely heterogeneous systems. MCL supports the execution of several multi-threaded applications within the same compute node, performs asynchronous execution of application tasks, efficiently balances computation across hardware resources, and provides performance portability. We show that code developed on a personal desktop automatically scales up to fully utilize powerful workstations with 8 GPUs and down to power-efficient embedded systems. MCL provides up to 17.5x speedup over OpenCL on NVIDIA DGX-1 systems and up to 1.88x speedup on single-GPU systems. In multi-application workloads, MCL dynamically resource allocation provides up to 2.43x performance improvement over manual, static allocation of computing resources.

Gioiosa, Roberto↗

Software defined grid energy storage

Today, consumer battery installations are isolated, physical devices. Virtual power plants (VPPs) allow consumer devices to aggregate for grid services, but they are are vertically integrated, vendor controlled systems (e.g., Tesla’s VPP). Consumer batteries are therefore unable to participate in energy markets or other grid services outside what their vendor provides. We describe a software system that provides software control of multiple, networked battery energy storage systems in the electric grid. The system introduces two new ideas that enable flexible and dependable management of energy storage. The first is a virtual battery, which can either partition a battery or aggregate multiple batteries. The second is a reservation-based API which allows asynchronous control of batteries to meet contractual guarantees in a safe and dependable manner. Virtual batteries and a reservation-based API address the unique challenges of achieving high and efficient utilization of energy storage systems, including heterogeneity of battery systems such as varying C-rates, participation in energy markets, utility bill management systems, community resource sharing, and reliability. Using a testbed comprised of sonnen Inc. storage units installed in several homes and a lab, we demonstrate that virtualized batteries can seamlessly replace physical batteries, flexibly manage energy storage resources, isolate multiple clients using a shared battery, and create new energy storage applications.

25 ENERGY STORAGE↗

Large-Scale Materials Modeling at Quantum Accuracy: Ab Initio Simulations of Quasicrystals and Interacting Extended Defects in Metallic Alloys

Ab initio electronic-structure has remained dichotomous between achievable accuracy and length-scale. Quantum many-body (QMB) methods realize quantum accuracy but fail to scale. Density functional theory (DFT) scales favorably but remains far from quantum accuracy. We present a framework that breaks this dichotomy by use of three interconnected modules: (i) invDFT: a methodological advance in inverse DFT linking QMB methods to DFT; (ii) MLXC: a machine-learned density functional trained with invDFT data, commensurate with quantum accuracy; (iii) DFT-FE-MLXC: an adaptive higher-order spectral finite-element (FE) based DFT implementation that integrates MLXC with efficient solver strategies and HPC innovations in FE-specific dense linear algebra, mixed-precision algorithms, and asynchronous compute-communication. Furthermore, we demonstrate a paradigm shift in DFT that not only provides an accuracy commensurate with QMB methods in ground-state energies, but also attains an unprecedented performance of 659.7 PFLOPS (43.1% peak FP64 performance) on 619,124 electrons using 8,000 GPU nodes of Frontier supercomputer.

density functional theory↗

Q-IRIS: The Evolution of the IRIS Task-Based Runtime to Enable Classical-Quantum Workflows

Extreme heterogeneity in emerging HPC systems are starting to include quantum accelerators, motivating runtimes that can coordinate between classical and quantum workloads. We present a proof-of-concept hybrid execution framework integrating the IRIS asynchronous task-based runtime with the XACC quantum programming framework via the Quantum Intermediate Representation Execution Engine (QIR-EE). IRIS orchestrates multiple programs written in the quantum intermediate representation (QIR) across heterogeneous backends (including multiple quantum simulators), enabling concurrent execution of classical and quantum tasks. Although not a performance study, we report measurable outcomes through the successful asynchronous scheduling and execution of multiple quantum workloads. To illustrate practical runtime implications, we decompose a four-qubit circuit into smaller subcircuits through a process known as quantum circuit cutting, reducing per-task quantum simulation load and demonstrating how task granularity can improve simulator throughput and reduce queueing behavior -- effects directly relevant to early quantum hardware environments. We conclude by outlining key challenges for scaling hybrid runtimes, including coordinated scheduling, classical-quantum interaction management, and support for diverse backend resources in heterogeneous systems.

Miniskar, Narasinga Rao [ORNL] (ORCID:000000018259↗

Integrating ytopt and libEnsemble to autotune OpenMC

Ytopt is a Python machine-learning-based autotuning software package developed within the ECP PROTEAS-TUNE project. The ytopt software adopts an asynchronous search framework that consists of sampling a small number of input parameter configurations and progressively fitting a surrogate model over the input-output space until exhausting the user-defined maximum number of evaluations or the wall-clock time. libEnsemble is a Python toolkit for coordinating workflows of asynchronous and dynamic ensembles of calculations across massively parallel resources developed within the ECP PETSc/TAO project. libEnsemble helps users take advantage of massively parallel resources to solve design, decision, and inference problems and expands the class of problems that can benefit from increased parallelism. In this paper we present our methodology and framework to integrate ytopt and libEnsemble to take advantage of massively parallel resources to accelerate the autotuning process. Specifically, we focus on using the proposed framework to autotune the ECP ExaSMR application OpenMC, an open source Monte Carlo particle transport code. OpenMC has seven tunable parameters some of which have large ranges such as the number of particles in-flight, which is in the range of 100,000 to 8 million, with its default setting of 1 million. Setting the proper combination of these parameter values to achieve the best performance is extremely time-consuming. Therefore, we apply the proposed framework to autotune the MPI/OpenMP offload version of OpenMC based on a user-defined metric such as the figure of merit (FoM) (particles/s) or energy efficiency energy-delay product (EDP) on Crusher at Oak Ridge Leadership Computing Facility. In conclusion, the experimental results show that we achieve the improvement up to 29.49% in FoM and up to 30.44% in EDP.

Autotuning↗

APOLLO: a facility-scale differentiable virtual accelerator at Fermilab FAST/IOTA

As the design complexity of modern accelerators grows, there is more interest in using advanced simulations that have fast execution time or yield additional insights like gradients. The FAST/IOTA facility has been working on implementing and experimentally validating an end-to-end digital twin that is both fast and gradient-aware, allowing for rapid prototyping of new software and experiments with minimal beam time costs. Our framework integrates physics and ML codes for linac and ring simulation through a set of generic interfaces between surrogate and physics-based sections. To reproduce device inputs and outputs, system state is exposed as a deterministic event loop in a specialized discrete event simulator architecture. Because Fermilab is undergoing control system transition, several APIs were implemented as final user interfaces - a fully asynchronous EPICS soft IOC, a gRPC-based Data Pool Manager (DPM), and legacy ACNET protocols. We discuss implementation details as well as challenges handling live data assimilation and future plans to extend modelling to main complex proton accelerators like PIPII and Booster.

Kuklev, Nikita [Fermilab]↗

Multi-Port Autonomous Reconfigurable Solar Power Plant

As the penetration level of power electronics increases and remote photovoltaic (PV) generation is integrated into the alternating current (ac) grid, the short-circuit ratio (SCR) at the point of interconnection of a hybrid PV–energy storage system (ESS) plant may be low. Additionally, the inertia of the alternating current (ac) grid may be low. The low SCRs and inertias can lead to reliability challenges in the power grid. These operating conditions require additional reinforcements, such as synchronous condensers, static var compensators, static synchronous compensators, and high-voltage direct current (HVdc) links/grids. HVdc links or grids may also provide the additional capability of connecting the plant to asynchronous grids and/or connecting asynchronous grids, among others. This scenario leads to discrete development of solar inverters, energy storage inverters, HVdc converters, and several transformers. Some of the problems associated with this discrete development and inverter-based generation include increased cost, lower reliability, and reduced efficiency associated with duplication of power electronics (PEs); competing controls of several individual discrete inverter-based generators due to the presence of multiple PEs, which leads to derating of the system; and transient stability problems arising from inverter-based generation, such as voltage/frequency events leading inverter shutdowns, voltage instability and control interactions in the formed weak grid, and harmonics caused by resonances of multiple inverters.

14 SOLAR ENERGY↗

Accelerating shared file checkpoint with local burst buffers

A data management system and method for accelerating shared file checkpointing. Written application data is aggregated in an application data file created in a local burst buffer memory at a compute node, and an associated data mapping built index to maintain information related to the offsets into a shared file at which segments of the application data is to be stored in a parallel file system, and where in the buffer those segments are located. The node asynchronously transfers a data file containing the application data and the associated data mapping index to a file server for shared file storage. The data management system and method further accelerates shared file checkpointing in which a shared file, together with a map file that specifies how the shared file is to be distributed, is asynchronously transferred to local burst buffer memories at the nodes to accelerate reading of the shared file.

Gooding, Thomas↗