Engineering PapersSearch

SEARCH · Engineering Papers

Results for “Load balancing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

UMap: An application-oriented user level memory mapping library

Exploiting the prominent role of complex memories in exascale node architecture, the UMap page fault handler offers new capabilities to access large memory-mapped data sets directly. UMap provides flexible configuration options to customize page handling to each application, including analysis of massive observational and simulation data sets. The high-performance design features I/O decoupling, dynamic load balancing, and application-level controls. Page faults triggered by application threads and processes accessing data mapped to a UMapp’ed region are handled via the Linux userfaultfd protocol, an asynchronous message-oriented kernel-user communication mechanism that avoids the context switch penalty of traditional signal fault handlers. UMap is fully open source. In this paper, we give an overview of the UMap library architecture, its extensible plugin architecture, and the use/performance of UMap in emerging heterogeneous memory hierarchies such as near-node Non-volatile Memory (NVM) and network attached memories. We highlight new capabilities in two pagefault management plugins, the NetworkStore and SparseStore. We demonstrate the integration between UMap and multiple ECP products including Caliper, Metall, ZFP, Mochi, and Ripples.

97 MATHEMATICS AND COMPUTING

Deployment of inference as a service at the US CMS Tier-2 data centers

Coprocessors, especially GPUs, will be a vital ingredient of data production workflows at the HL-LHC. At CMS, the GPU-as-a-service approach for production workflows is implemented by the SONIC project (Services for Optimized Network Inference on Coprocessors). SONIC provides a mechanism for outsourcing computationally demanding algorithms, such as neural network inference, to remote servers, where requests from multiple clients are intelligently distributed across multiple GPUs by a load-balancing service. This talk highlights the recent progress in deploying SONIC at selected U.S. CMS Tier-2 data centers. Using realistic CMS Run3 data processing workflows, such as those containing transformer-based algorithms, we demonstrate how SONIC is integrated into the production-like environment to enable accelerated inference offloading. We will present developments from both the client and server sides, including production job and data center configurations for NVIDIA and AMD GPUs. We will also present performance scaling benchmarks and discuss the challenges of operating SONIC in CMS production, such as server discovery, GPU saturation, fallback server logic, etc.

Holzman, Burt

Exploring DAOS as a Burst Buffer for a 100 Gbps DAQ Real-Time Streaming System

We present an experimental evaluation of a burst buffer for a real-time DAQ streaming system designed to transmit instrument data to remote data centers. The system is based on EJ-FAT, a load balancing system capable of Nx 100Gbps streams, distributing data from event sources to processing nodes. We explore applying the DAOS system as a burst buffer to serve a number of purposes: improve resiliency, elasticity and add new functions into the processing pipeline. In the evaluation a sender transmits events over a 100Gbps network to a receiver integrated with DAOS to store the reassembled events using DAOS APIs. We evaluate the system for possible bottlenecks and provide end-to-end evaluation with a burst buffer using DAOS storage abstractions. We show that a receiver node can support 38.1 Gbps. This proves the viability of our approach and allows us to extend this work to investigate scale-out properties and new streaming optimizations.

Mei, Xinxin

An Efficient Storage-Driven Machine Learning Model for Performance in the Era of Multimodal Scientific Data

Scientific workflows are increasingly relying on machine learning (ML), simulation, and hybrid techniques to predict, understand, and optimize the behavior of complex experiments. High-performance computing has greatly improved researchers’ ability to acquire diverse data modalities in these workflows. Recent studies suggest that the performance of machine learning models can be improved by integrating data from various sources. Unfortunately, these workloads pose unprecedent pressure on the network storage to meet the demands associated with accessing these multimodal data. To mitigate the impact of intensive IO, we propose a solution that utilizes a multi-tier High-Performance Computing (HPC) distributed storage and data processing framework, placing computation where the data resides for better performance. By adopting this project, the scientific community will gain new opportunities to explore multimodal storage-driven possibilities, integrating multiple scientific data sources with advanced streaming frameworks. Additionally, our framework effectively utilizes computing resources and bridges the gaps identified by HPC experts. Our proposed approach tackles scalability and persistence challenges by leveraging native persistency, which has posed difficulties in traditional approaches. Furthermore, we seek to enhance fault-tolerance and load-balance of computations by leveraging real-time streaming in diverse scientific computing environments, thereby propelling advanced scientific computing research into the next generation.

97 MATHEMATICS AND COMPUTING

Advancing Grid Resilience through Smart Charge Management: Findings from Maryland’s Pilot

This report presents research findings from a four-year Smart Charge Management (SCM) pilot program conducted by Maryland’s largest electric utilities—Baltimore Gas and Electric (BGE), Potomac Electric Power Company (Pepco), and Delmarva Power & Light (DPL)—to evaluate strategies for optimizing electric vehicle (EV) charging loads and enhancing grid stability. Supported by the U.S. Department of Energy (DOE), Argonne National Laboratory collaborated with all project partners and examined the effectiveness of Time-of-Use (TOU) and Load Balancing (LB) strategies in managing peak demand, deferring costly infrastructure upgrades, and reducing grid constraints at the feeder level.

24 POWER TRANSMISSION AND DISTRIBUTION

EJFAT Scientific Perspective

Presented new computing model to the test by deploying the EJFAT system alongside a data-stream processing framework running the production-level CLAS12 event reconstruction application. In this experiment, a continuous stream of CLAS12 Level-1 identified events was processed in real-time using the EJFAT load balancer, distributing the workload across 90 computing nodes located across the U.S. This marks the first-ever large-scale, real-time distributed data stream processing experiment, demonstrating that scientific data-streaming pipelines can efficiently scale across four dimensions, thanks to EJFAT’s advanced hardware and software capabilities.

Gyurjyan, Vardan [Thomas Jefferson National Accele

Synchronous Machine Governor Upgrade

Conventional generation sources play a critical role in the stability and reliability of the electrical grid, particularly as we transition towards more renewable energy sources. To understand and accurately emulate their behavior for optimizing grid operations and ensuring seamless integration with renewable technologies, it is essential to better emulate the grid- and plant-level impacts of conventional generation sources, such as natural gas (NG) driven heat recovery steam generators (HRSGs) and combustion turbines (CTs). Therefore, a governor model is developed in a programmable logic controller (PLC) to investigate the performance of the conventional generator under various dynamic operating conditions and to identify the impact on grid stability in a controlled environment. The governor model aims to enable the hardware-in-the-loop (HIL) based emulation of these conventional generation sources using the existing 2 MVA synchronous machine/generator that is driven by a flexible 2.5 MW variable speed drive. This setup will allow us to replicate the dynamic characteristics and response behaviors of NG-driven HRSGs and CTs. The controls for the emulated conventional plants follow the industry standard and are adjustable, ensuring they accurately reflect the operational capabilities and limitations of real-world systems. These controls include load-following capabilities, ramp rates, startup and shutdown sequences, and emissions characteristics. By incorporating these adjustable controls, we aim to capture the nuanced impacts of conventional generation, such as their ability to provide ancillary services like frequency regulation, voltage support, and spinning reserve. In this report, we simulate two types of dynamic operations: grid-connected and islanding. For each dynamic operation, representative starting sequences are tested, including turbine purge, ignition, speed ramping up, generator excitation and synchronizing, and breaker close. The HIL based tests provides insights for field deployment, specifically the high-fidelity governor model provides results to predict the potential stability and reliability risk and suggest possible integration measures (e.g., generation and load balancing, tuning of governor control parameters). Ultimately, this enhanced emulation capability will be integrated into our Advanced Research on Integrated Energy Systems (ARIES), enabling us to conduct comprehensive studies on the interactions between conventional and renewable energy sources. By better understanding these interactions, we can develop strategies to optimize the overall performance and reliability of the grid. This will support the deployment of advanced grid management techniques, such as demand response, grid-forming inverters, and energy storage systems. The main contributions are summarized as follows: (1) This report introduces a PLC-based governor model for gas turbines. This model accurately simulates the dynamic behavior of conventional generation sources under various operational scenarios; (2) The model is integrated with an HIL testbed that includes a 2.5 MW variable speed drive and a 2 MVA synchronous machine. This setup enables realistic, real-time emulation of conventional power plants, particularly NG driven HRSGs and CTs; (3) The developed model is adaptable to various gas turbine configurations and allows for precise control over parameters such as MW ramp rates. This flexibility makes it a valuable tool for future research and industry collaboration; and (4) By incorporating the model into the National Renewable Energy Laboratory's Advanced Research on Integrated Energy Systems, the report lays the groundwork for future studies on interactions between conventional and renewable energy sources, enhancing the ability to develop advanced grid management strategies.

24 POWER TRANSMISSION AND DISTRIBUTION

DUNE Rucio Server Scalabiilty Studies

The DUNE collaboration has an ongoing production effort to simulate the full detectors and to analyze the various prototypes that are currently running. Rucio is used to manage the 40PB of files made to date. When 500 or more jobs were sending output to Rucio simultaneously via Rucio upload, we observed timeouts, unhandled exceptions, and Rucio server restarts due to slow performance. In collaboration with the core Rucio team we did a full review of the Rucio upload code and identified several optimizations that can be made. We also have deployed the Ingress load balancer in front of our Rucio servers and added a database connection pooling utility. These changes led to significant improvement both in reliability and scalability, yet we anticipate even better performance will eventually be required. We describe in this paper the initial state of the system, the various debugging processes that were used, and our plans to further improve scalability.

Calcutt, J. [Brookhaven Natl. Lab.]

Peer-to-peer communication control for resilient operations of networked cyberphysical systems

This report includes two main accomplishments of the peer-to-peer communication control for resilient operation of networked microgrids project in FY24, which include a scheme for cyberattack-aware coordination of networked microgrids for supporting voltages of bulk power systems and a scheme for price signal-based operations of EV-rich networked microgrids with mixed ownership. First, the cyberattack-aware scheme enables networked microgrids to distributedly determine the amount of reactive power injection to support the voltage of bulk power system (BPS) in a fair manner. In this scheme, a risk-informed algorithm is presented to generate the peer-to- peer (P2P) communication graph with minimal risk of attack on communication links. To deal with cyberattacks on MG controllers, the resilient consensus algorithm (CA) is utilized for MG controllers to robustly estimate the total reactive power headroom, from which the MGs can accurately provide the needed amount of reactive power injection for supporting the voltage of BPS. The CA implementation and performance within the P2P communication framework are demonstrated on the IEEE 39-bus system with 6 microgrids contained in the distribution feeder under different cyberattack scenarios. Second, the price-based scheme enables the usage of the real-time price signal for the operations of electric vehicle (EV)-rich networked-microgrids with mixed ownership, in which not all the microgrids can communicate with the distribution system operator (DSO). In this scheme, a max consensus is introduced to enable the real-time price signal to be propagated from the DSO to all the microgrids, from which each microgrid controller will manage the DERs to balance the load demand and the power injection from the EV charging stations within its microgrid. Numerical results over one day with 288 slots of 5-minute intervals on the modified 123-node test feeder including 3 microgrids with high penetration of EV are presented to evaluate how the price signal affects the operations of networked microgrids under different charging strategies of the EV charging stations. The result indicates that our proposed EVCS (dis)charging strategy, which leverages the flexibility of EVs to support the grid through discharging during peak demand, proves to be a cost-effective solution that reduces operational costs while improving the social welfare of EV charging.

24 POWER TRANSMISSION AND DISTRIBUTION

Virtual Power Plants and Distributed Energy Resource Management Systems

Virtual Power Plants (VPPs) are aggregations of DERs that can balance electrical loads and provide utility-scale and utility-grade grid services like a traditional power plant. This presentation covers VPP definition, State-of-the-Art, Grid Architectures, Example VPP studies, VPP Standards, and VPP Roadmap.

24 POWER TRANSMISSION AND DISTRIBUTION

Design of a SMART Valve Testbed for Nuclear Thermal Dispatch

By the year 2050, the United States aims to achieve net-zero carbon emissions. To achieve this target, the licensing of the Light Water Reactor (LWR) fleet has been extended for 20 more years. To stay economically competitive with other power sources such as renewable and fossil-fuel power plants, the U.S. Department of Energy has introduced a plan to modernize the existing LWR fleet and diversify the revenue stream. One of the plans is to dispatch thermal energy to endothermic industrial processes. SMART valves will play an important role in this initiative by efficiently balancing the load by regulating valves in a coordinated manner while monitoring the thermal-hydraulic systems to enhance safety and maintain the integrity of the power plant. This research aims to develop a facility to test the coordinated control algorithm and produce various test results for training the monitoring system. The constructed facility is capable of simulating various operational and accidental scenarios by coordinating all the valves (positions) and pump (flowrate). The facility is developed with an Internet of Things (IoT)-based custom system and a python-based valve position control and coordination mechanism. It has achieved stable sensor outputs, pump control, and coordinated valve regulation in all three valves with minimum obstruction in the system.

22 - GENERAL STUDIES OF NUCLEAR REACTORS

Solving the Grid Optimization Competition Challenge 3 Problem

The Grid Optimization Competition Challenge 3 Problem posed a multiperiod security-constrained unit commitment problem with base-case AC power flow. The problem formulation includes binary unit commitment decisions, nonlinear AC power flow and balance, dispatchable loads, and linearized contingency real power flow, among other features. This talk will present a modified consensus ADMM algorithm, which splits the problem into mixed-integer linear and nonlinear components, as a heuristic solution method for this large-scale mixed integer nonlinear program. We will present some computational results from the competition for our implementation and reflect on the challenges of participating the grid optimization competition.

AC power flow

A Hardware-in-the-Loop Experimental Testbed Using Air Conditioners for Grid Balancing

Driven by the need to offset the variability of renewable generation on the grid, development of load control is a highly active field of research. However, practical use of residential loads for grid balancing remains rare, in part due to the cost of communicating with large numbers of small loads and also the limited experimentation done so far to demonstrate reliable operation. To establish a basis for the safe and reliable use of fleets of compressor loads as distributed energy resources, we constructed an experimental testbed in a laboratory, so that load coordination schemes could be tested at extreme conditions. Here, this experimental testbed was used to tune a simulation testbed to which it was then linked, thereby augmenting the effective size of the fleet. Modeling of the system was done both to demonstrate the experimental testbed's behavior and also to understand how to tune the behavior of each load. Implementing this testbed has enabled rapid turnaround of experiments on various load control algorithms, and year-round testing without the constraints and limitations arising in seasonal field tests with real houses. Experimental results show the practical feasibility of an ensemble of small loads contributing to grid balancing.

Air Conditioners

Software Control Program For Transportable Microgrid State-of-charge Balancing And Frequency Stability Controls

A deterministic state-of-charge (SOC) balancing approach software control code is introduced as an integral secondary management to primary control layer of an islanded small microgrid or nanogrid system made up of multiple grid-forming inverter/battery/solar combination systems, where each set of batteries with each inverter are on independent DC buses (i.e. non-paralleled on the DC sides). A DERMS-level control approach, algorithm and automation controller program was developed to improve coordination and enable microgrid asset compliance and SOC balancing, enabling provision of a system-level power stability support architecture, load support, and asset scalability. The architecture is configured to treat each unit or micro/nano-grid as a node in a microgrid network, allowing for autonomous DERMS control regarding load and SOC balancing and power stability. As the network grows with the addition of units, greater coordination efforts may be required. The ideal small network microgrid ranges from 2-10 inverter/battery units before additional control parameters must be considered in the existing architecture. The control approach focuses on a deterministic state-of-charge analysis as the primary level control process followed by a secondary control loop using a forced frequency-watt droop strategy to conform off-the-shelf components into behaving under a leader-follower configuration. Adopting this control scheme has been shown to allow for a balanced, unit-coordinated microgrid network, enabling stable power flow. The deterministic state-of-charge approach is introduced as an integral primary control layer of an islanded small network microgrid. A standard strategy for SOC balancing is implementing a battery management system (BMS) to control SOC on the DC side. An alternative approach is to determine how to coordinate sending and receiving power on the AC side with multiple units. The latter approach assesses all the integrated units in the microgrid network. Once the individual units are identified, further system data is required to calculate each unit's total kWh, provided information about its capability to supply or consume kWh and availability. The secondary control layer in the multi-layered small network microgrid methodology uses the primary layer’s decision to initiate frequency setpoint changes, initializing the SOC balancing. The secondary control layer considers numerous system-dependent variables to enable a charging and discharging profile based on adjustable frequency setpoints. The combined architecture will result in stable, coordinated power flow enhancing an AC microgrid's functionalities.

Myers, KurtS [Idaho National Laboratory (INL), Ida

Hourly Electricity Demand Projections for Eight Combined Climate and Socioeconomic Scenarios

This dataset contains 40 years (1980-2019) of simulated historical hourly electricity demand (i.e., loads) and 80 years (2020-2099) of projected hourly loads for 54 Balancing Authorities (BAs) and 48 states plus the District of Columbia. Details about the scenarios and variables included in this dataset are in the readme.pdf file. The two primary models that created the dataset are a version of the Global Change Analysis Model with detailed sectoral resolution over the United States (GCAM-USA) and the Total ELectricity Loads (TELL) model. Links to the model source code and workflow for deriving the dataset are provided in an accompanying meta-repository: https://github.com/IMMM-SFA/burleyson-etal_2024_applied_energy. Projections are for four future climate scenarios that represent combinations of Representative Concentration Pathways (RCPs) 4.5 and 8.5 combined with two levels of climate model sensitivities: rcp45cooler, rcp45hotter, rcp85cooler, and rcp85hotter. The four climate scenarios are crossed with Shared Socioeconomic Pathways (SSPs) 3 and 5 to yield eight different future load projections: rcp45cooler_ssp3, rcp45cooler_ssp5, rcp45hotter_ssp3, rcp45hotter_ssp5, rcp85cooler_ssp3, rcp85cooler_ssp5, rcp85hotter_ssp3, and rcp85hotter_ssp5. The climate scenarios are from the IM3 Thermodynamic Global Warming (TGW) dataset which is linked below in the related metadata. The related metadata also contains links to a repository containing the raw GCAM-USA output files.

Climate Change

Scalable Solution-Processed Electrolyte Membranes with Optimized Microstructure for High-Performance Protonic Ceramic Electrochemical Cells

Proton-conducting electrochemical cells (PCECs) are promising for efficient hydrogen production, but achieving dense, uniform, thin electrolyte layers remains a key challenge, particularly for scalable fabrication. Here, we present a solution-processed deposition approach with a mechanistically optimized slurry for uniform electrolyte formation. By tailoring particle size distribution, solid loading, and solvent/additive balance, we regulated wetting behavior and evaporation kinetics of the electrolyte slurry to promote homogeneous electrolyte particle packing. These features facilitate tight grain boundary contact and early stage neck growth during sintering, eliminating residual porosity, and improving mechanical integrity. The resulting ∼15 μm thick electrolyte shows high density, strong electrode adhesion, and stable interfaces outperforming the previously reported spray-based fabricated electrolyte by about 31% at 600 °C in FC mode. Single cells deliver 0.962 W cm –2 at 600 °C in fuel cell mode and 1.31 A cm –2 at 1.3 V in electrolysis mode, maintaining robust performance over 100 h with negligible degradation (≤0.02% h –1 ) in each mode. Scale-up to 2.5 cm diameter substrates confirmed reproducible densification and geometric stability. This work demonstrates a cost-effective, scalable route where control over particle-fluid interactions and drying dynamics enables a superior electrolyte microstructure and high PCEC performance.

dense microstructure

Maximizing long-term biohydrogen production with Clostridium thermocellum for high solids conversion of lignocellulosic biomass

Biological hydrogen production from lignocellulosic biomass sustainably couples organic waste reduction with renewable energy generation. Efficient conversion is challenged by the structural complexity of lignocellulose and resulting recalcitrance to enzymatic degradation. Clostridium thermocellum natively breaks down biomass with highly effective hemi-/cellulases systems (i.e., cellulosomes) and generates hydrogen in anaerobic cultivation, creating a compelling platform for lignocellulosic biohydrogen production. Achieving commercially viable production rates requires balancing high biomass loading and throughput against uniform mixing conditions required for enzyme dispersion, pH and temperature control, and efficient hydrogen and metabolite removal in continuous operation. To address these barriers to process intensification, we implemented novel reactor and process designs for high-solids lignocellulosic biomass fermentations using the C. thermocellum KJC19-9 strain, genetically engineered for co-utilization of cellulose and hemicellulose sugars (i.e., xylose). Via computational fluid dynamics (CFD) modeling and experimental validation, we achieved a >50% improvement in biohydrogen production with an improved anchor-type impeller morphology, coupled to a threefold reduction in agitation rate. To further reduce rheological constraints and accumulation of toxic metabolites, we then transitioned the process to sequencing fed-batch operation. The resulting process generated 24.87 L H 2 L −1 from 160 g L −1 of deacetylated and mechanically refined (DMR)-pretreated corn stover biomass over 16 days while solubilizing >95% of influent cellulose and hemicellulose, setting a new performance benchmark for continuous production of biohydrogen from lignocellulose.

08 HYDROGEN

A Trade-Off Study Between the Primary and Transient Responses of Grid-Forming Inverters

The control parameters of the grid-forming (GFM) inverter-based resources (IBRs) directly impact power system dynamics. The primary control requires the GFM inverter to balance generation and load. It is also preferred that a GFM inverter maintain the terminal frequency after a power system fault. In this paper, we investigate the trade-off between the primary control objective and the transient response. To quantify the transient performance of the GFM inverter, we introduce a new real-time transient stability index (TSI). This index plays a crucial role in our investigation, as it allows us to compare the performance trade-offs of different sets of GFM control parameters.

Lin, Xuheng