Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Cascading Failure”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

Complete Development of Critical Capabilities for TRISO Fission Product Source Term Calculations and Quantify Mechanisms for Pd Penetration of SiC

Overall fission product (FP) release will be an important consideration for the licensing and deployment of advanced reactors utilizing tristructural isotropic (TRISO) fuels. This work focuses on enhancing and applying the BISON models needed to predict FP transport within TRISO particles and particle failure probability, both of which factor directly into release predictions. Specifically, this report details (1) the development of the models needed to predict palladium (Pd) conservation at the engineering scale and the application of those models to characterize Pd fluxes for input into a mechanistic multiscale model for Pd penetration; (2) the refinement of sorption mass transfer models and the development of models for trapping in porous layers, which were applied and compared to particle scans from AGR-2 to provide proof of concept for a method of particle-scale validation that may reduce uncertainties compared to compact-scale validation using data from integral effects tests; (3) the development of a failure-statistics-informed, mesh-independent methodology for applying smeared cracking, enabling further study of the localized multiphysics behaviors associated with cascading particle failure mechanisms; and (4) the preliminary characterization of those coupled multiphysics particle failure behaviors using smeared, nonretentive diffusivities to provide a baseline for future study and to guide ongoing engineering applications.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Efficient Contingency Analysis in Power Systems via Network Trigger Nodes

Modeling failure dynamics within a power system is a complex and challenging process due to multiple inter-dependencies and convoluted inter-domain relationships. Subject matter experts (SMEs) are interested in understanding these failure dynamics for reducing the impact from future disasters (i.e., losses or failures of power system components, such as transmission lines). Contingency analysis (CA) tools enable such ’what-if’ scenario analyses to evaluate the impacts on the power system. Analyzing all possible contingencies among N system components can be computationally expensive. An important step for performing CA is identifying a set of k ‘trigger’ components, which when failed initially can significantly impact the overall system by causing multiple failures. Currently SMEs focus on identifying these trigger components by running expensive simulations on all possible subsets, which quickly becomes infeasible. Hence finding a relevant set of trigger components (contingencies) rapidly to enable efficient and useful CA is crucial.In a collaboration between computer scientists and power system experts, we propose an efficient method for performing CA by exploiting network inter-dependencies in power system components. First, we construct a network with multiple electric grid infrastructure components and dependencies as connections among them. We reformulate the problem of finding a set of trigger components as a problem of identifying critical nodes in the network, which can cascade power failures through connected nodes and cause significant damage to the network. To guide the practical CA tools, we develop a network-based model with a probabilistic edge-weights setup using intricate domain rules. Then we conduct an empirical study on real power system data in the US for both regional and national levels. Firstly, we use power system datasets for the US to create a national-scale domain-driven model. Secondly, we demonstrate that network-based model outperforms the outputs from a real CA tool and show on average 25 × improved selection of contingencies, thereby showcasing practical benefits to the power experts.

Tabassum, Anika↗

How extreme rainfall and failing dams unleashed the Derna flood disaster

On September 11, 2023, Storm Daniel unleashed unprecedented rainfall over the Wadi Derna watershed, triggering one of the most devastating floods in modern history, striking Derna, a coastal city in Libya. This study reconstructs the disaster using an integrated modeling approach that combines satellite imagery, hydrologic, hydraulic, and geotechnical simulations, machine learning, eyewitness accounts, and digital elevation data to assess the impact of cascading dam failures. Our findings reveal that the region’s dams, even if structurally sound, would have provided minimal protection against the extreme runoff. However, their failure unleashed a destructive surge wave, amplifying the disaster’s magnitude and devastation. Here, we show that the collapse of aging flood control infrastructures, compounded by inadequate risk assessment and emergency preparedness, dramatically escalated the disaster’s impact. Our findings underscore the urgent need for systematic dam safety evaluations, enhanced flood forecasting, and adaptive risk management strategies that address climate extremes and infrastructure vulnerabilities.

Hydrology↗

Mitigation of Motor Stalling and FIDVR via Energy Storage Systems with Temporal Logic Specifications

The fault-induced delayed voltage recovery (FIDVR) phenomenon has been very common from the distribution system through the transmission system. It causes a delay on recovering significantly depressed local voltage after the fault is cleared, and it can also lead to more widespread cascading system failures. Mitigating this event with current control approaches is challenging and becoming a crucial issue. Here, a model predictive control-based strategy employing signal temporal logic specifications is proposed to help mitigate FIDVR. To this end, it investigates and extends a dynamic performance model allowing analytic insights into the system-wide impact of motor stalling and FIDVR. The proposed controller provides richer descriptions of voltage specifications addressing both magnitude and time simultaneously. We consider different control specifications with reactive power support from energy storage systems to prevent the voltage during/after the fault from dropping too low, and reduce the delay time of voltage recovery. The simulation results conducted with the IEEE 57 bus test network validate the proposed method and demonstrate the effectiveness of the mitigation strategy on FIDVR.

25 ENERGY STORAGE↗

A Comprehensive Numerical and Experimental Study for the Passive Thermal Management in Battery Modules and Packs

Cooling plates in battery packs of electric vehicles play critical roles in passive thermal management systems to reduce risks of catastrophic thermal runaway. In this work, a series of numerical simulations and experiments are carried out to unveil the role of cooling plates (both between cells and a bottom plate parallel to the cell stack) on the thermal behavior of battery modules and packs under nail penetrations. First, we investigated the role of side cooling plates on the thermal runaway propagation mitigation in battery modules (1S3P) and packs (3S3P) by varying the key parameters of the side cooling plates, such as plate thicknesses, thermal contact resistances, and materials. Then, three important factors for passive thermal management systems are identified: (i) thermal mass of side cooling plates, (ii) interfacial thermal contact resistances, and (iii) the effective heat transfer coefficients at exterior surfaces. The roles of bottom cooling plates on thermal runaway propagation mitigation in 1S3P and 1S5P battery modules are numerically investigated by comparing the thermal behavior of the modules with only side cooling plates and with both side and bottom cooling plates.

25 ENERGY STORAGE↗

An Improved Transmission Switching Algorithm for Managing Post-(N-1) Contingencies in Electricity Networks

In this presentation, we detail the shortcomings of existing transmission switching (TS) methodologies for preventing post-contingency loss of load in wide area electric power networks. Following that, we present a bi-level algorithm as an improvement to the state of the art. We show proof-of-concept level results that indicate computational viability, fast solutions, and the ability to pick the same best candidate as non computationally viable methods. Further, our method also minimizes load shedding after the contingency, thus enhancing the reliability and security of electricity supply. We also include a futuristic scenario of increased renewables and decreased natural gas sourced generators to show promising results. We grant permission for our presentation to be recorded and uploaded by the conference organizers.

24 POWER TRANSMISSION AND DISTRIBUTION↗

A Computationally Improved Heuristic Algorithm for Transmission Switching Using Line Flow Thresholds for Load Shed Reduction

We present a computationally improved heuristic algorithm for transmission switching (TS) to recover load shed. Research from the past showed that changing power system topology may control power flows and remove line congestion. Hence, TS may reduce the required load shed. One of the main challenges is to find a potential TS candidate in a suitable time. Here, we propose a novel heuristic method that is capable of finding the potential TS candidate faster than existing algorithms in literature. The proposed method is compatible with both the AC and DC optimal power flows (OPF). Three metrics are used to compare the proposed algorithm with the state-of-the-art from literature to show the speedup and accuracy achieved. The proposed method is implemented on the IEEE 30-bus system, PEGASE 89-bus system, IEEE 118-bus system, and Polish 2383- bus system. The results on the large-scale Polish 2383-bus system shows that the proposed algorithm is scalable to large real-world systems. Parallel computing is implemented to further improve the computational performance of the proposed algorithm.

24 POWER TRANSMISSION AND DISTRIBUTION↗

AN INITIAL LOOK AT THE HEAT PIPE RESILIENCY TO REACTOR OPERATION

Several special purpose reactor designs are aiming to utilize heat pipes due to its inherently passive features allowing safe heat removal to the power conversion systems. There is insufficient data on heat pipe survivability in reactor environments, although heat-pipe failures are predicted to have low failure rates. In this paper, we perform a coupled neutronics and thermo-mechanics analyses for a 45 kWth HALEU fueled, hydride moderated homogenous design to evaluate the transient effects during both startup and a heatpipe-failure scenario. The heatpipe-failure study includes both a single failure case and a failure-propagation (cascading) case. While the model only evaluates a homogenous core, the approach could be applied to more complex models for any heat pipe cooled small reactor.

99 GENERAL AND MISCELLANEOUS↗

Data Centers and Digital Assurance Introduction to Supply Chain and Cybersecurity for Data Centers, Session 1

The first session of the TADA (Technical Assistance for Digital Assurance) Data Centers Cohort Workshop, held on October 30, 2025, introduced foundational concepts of Digital Assurance in the context of data center and grid integration. Sponsored by the U.S. Department of Energy, the workshop brought together utilities, data center operators, developers, and vendors to address cybersecurity and supply chain vulnerabilities. The session emphasized the growing criticality of data centers within the electric grid and the need for secure, real-time, bidirectional communication. Participants explored the principles of Digital Assurance, including cybersecurity, cyber-informed engineering (CIE), and lifecycle security, and applied a threat-vulnerability-consequence framework to identify and mitigate risks at the data center–grid interface. Discussions covered a range of threats such as spoofed dispatch signals and insider threats, architectural vulnerabilities like SCADA interfaces and insecure protocols, and potential consequences including cascading grid failures. The session also raised strategic questions about business value, vendor assurance, and defining cyber boundaries and responsibilities. This foundational workshop set the stage for deeper technical analysis and the development of actionable frameworks in subsequent sessions. Session 1 of 3.

24 - POWER TRANSMISSION AND DISTRIBUTION↗

Predictive Models and Novel Accelerated Tests for the Reliability of Cell Metallization and Solder Joints in Photovoltaic Modules

Interconnect, solder, and metallization-related failures are the primary reason for premature Si-based photovoltaic panel failure [1, 2] and can lead to a cascade of other failures, including thermal events. This combination of high severity and high occurrence means that understanding interconnect-related failures is key to the long-term health of the photovoltaic industry, and it is crucial that the industry adopt efficient and effective tests that aid in design-for-reliability and manufacturing quality control.

14 SOLAR ENERGY↗

Machine Learning and Data Science to Advance Laboratory Earthquake Prediction and Illuminate the Mechanics of Precursors to Failure

Earthquakes represent one of our greatest natural hazards and in recent years human induced seismicity is adding to the threat. Even a modest improvement in the ability to forecast devastating large earthquakes or smaller shallow events associated with fluid injection could save thousands of lives and billions of dollars. Current efforts to forecast earthquakes are limited by knowledge of earthquake physics and hampered by a lack of reliable lab or field observations. However, recent work has provided a critical opportunity for advancement. We have found: 1) clear and consistent precursors prior to earthquake-like failure in the laboratory and 2) that lab earthquakes can be predicted using machine learning (ML). These works show that stick-slip failure events –the lab equivalent of earthquakes– are preceded by a cascade of micro-failure events that radiate elastic energy in a manner that foretells catastrophic failure. Remarkably, ML predicts the fault zone stress state, the failure time and in some cases the magnitude of lab earthquakes. In addition, the observations include clear precursors to failure in the form of changes in fault zone properties prior to lab earthquakes. Precursors have been observed in previous laboratory studies but their origin is poorly understood and their possible connection to ML based earthquake prediction is unknown. The work conducted under our project has dramatically expanded these efforts. We have developed an integrated data science approach to illuminate the physics of earthquake precursors and lab earthquake prediction. Our work has accelerated the development of ML, artificial intelligence (AI), and related data science approaches by providing massive data sets that are tightly connected to critical scientific problems and by bringing together leading subject matter experts and data scientists. Earthquake physics involves phenomena that are far from equilibrium. Our work has leveraged data science methods to illuminate these phenomena and investigate how they relate to earthquake prediction. In addition to a large database with many types of labeled events that is available to everyone, our work has advanced the fundamental understanding of seismic forecasting, earthquake physics, and fault rheology

58 GEOSCIENCES↗

Machine Learning Approaches to Predicting Induced Seismicity and Imaging Geothermal Reservoir Properties

This project developed machine learning (ML) methods, lab data sets, and field data to advance geothermal exploration and geothermal energy production. The work had three focus areas. One involved the development of ML methods to use microearthquakes (MEQs) for imaging geothermal reservoir properties and improving subsurface characterization – most importantly the evolution of permeability within the evolving reservoir. This part of the work included development of ML approaches for automated MEQ location, focal mechanism determination and identification of earthquake precursors. The second area focused on using MEQ signals generated by geothermal exploration and production to predict the relationship between fluid injection and seismicity. Here, we extended to reservoir scale our success in using ML to predict laboratory earthquakes and fault zone stress state. The third focus area was on lab experiments. Here, we developed new ML models for lab earthquake prediction and identification of precursors to failure to improve earthquake forecasting and early warning in geothermal settings. Major outcomes of our work include ML models that learn from MEQ signals during geothermal exploration and production to predict induced seismicity. MEQs occur naturally in connection with drilling and energy production. We developed ML methods to use the seismic waves from these events to characterize the elastic, hydraulic and poromechanical properties of reservoirs. Our work illuminated fracture geometry and the evolution of fracture permeability by incorporating seismic coda wave analysis and ML methods to relate fluid injection and seismicity. We significantly expanded laboratory earthquake prediction to include methods that use both passive measurements of microearthquakes within the lab fault zones and also active source acoustic measurements of fault zone elastic properties. These methods can now predict fault zone stress state, time to failure and the magnitude of lab earthquakes. Our work showed that repetitive stick- slip failure events during frictional sliding (the lab equivalent of earthquakes) are preceded by a cascade of micro-failure events that radiate energy in a manner that foretells unstable failure – manifest as laboratory MEQs. We documented a mapping between fracture properties and statistical attributes of elastic radiation. We extended existing works to geothermal reservoir scale and developed ML methods to determine reservoir permeability, fracture properties, and their evolution during geothermal energy production. An attractive feature of ML algorithms is their ability to handle big datasets and reveal patterns and correlations that may remain invisible to conventional analyses. Our work connected data from field, laboratory and intermediate scales to study permeability, stress, strength, fracture stiffness and geometry. At the field scale we used data from the Newberry Volcano field site, UtahFORGE, EGS Collab, and also the Bedretto underground research lab in Switzerland. These data sets are bridging the gap between the lab scale, theory, and reservoir scale. Our work produced plain language summaries to improve public understanding of DOE research. We also developed openly distributed ML and seismicity datasets for use by all researchers and we published connections between induced seismicity in geothermal areas and reservoir properties including permeability, fracture properties, and stress state. Our models are designed for the large data sets of induced seismicity typically associated with geothermal sites. We produced labeled event catalogs and used them on geothermal data to assess how ML can facilitate geothermal production and exploration. All datasets are available on the GDR Productivity: The project produced 32 publications in peer reviewed journals (two are in review). It supported the work of 6 PhD students, 40 conference presentations, 6 keynote talks at national meetings, and mentoring and professional development for 4 postdoctoral fellows.

15 GEOTHERMAL ENERGY↗

A Hybrid Dynamic/Steady-State Tool With Protection Simulation for Cascading-Outage Analysis of Extreme Events in Power Systems

The bulk electric power grid is subject to vulnerabilities from component outages, which in certain combinations (extreme events) might lead to cascading outages. Some of these outages can be severe enough to trigger brownouts and blackouts. Much is known about mitigating the first few failures near the beginning of a cascade, but there are few established methods and tools for directly analyzing the risks of cascading component outages over a longer time scale. Current power system tools have limited ability to perform detailed and accurate cascading-outage analysis, which could be computationally intensive. The Dynamic Contingency Analysis Tool (DCAT) enables power system planning engineers to more realistically assess the consequences of extreme contingencies and potential cascading events across their systems and interconnections. DCAT has several unique features: (i) detailed hybrid dynamic and steady-state analysis of power systems to mimic real-world cascading outages, (ii) detailed modeling of protection systems embedded in the dynamic simulation, (iii) simulation of corrective action after transients, (iv) simulation of islanding , and (v) high-performance computing capability to simulate a large number of contingencies in a reasonable time. DCAT outputs will help find technically sound solutions to reduce the risk of cascading outages. This paper provides details of DCAT methodology and shows its capabilities with extreme events on real-world cases.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Adjustable Autonomy Testbed

The Adjustable Autonomy Testbed (AAT) is a simulation-based testbed located in the Intelligent Systems Laboratory in the Automation, Robotics and Simulation Division at NASA Johnson Space Center. The purpose of the testbed is to support evaluation and validation of prototypes of adjustable autonomous agent software for control and fault management for complex systems. The AA T project has developed prototype adjustable autonomous agent software and human interfaces for cooperative fault management. This software builds on current autonomous agent technology by altering the architecture, components and interfaces for effective teamwork between autonomous systems and human experts. Autonomous agents include a planner, flexible executive, low level control and deductive model-based fault isolation. Adjustable autonomy is intended to increase the flexibility and effectiveness of fault management with an autonomous system. The test domain for this work is control of advanced life support systems for habitats for planetary exploration. The CONFIG hybrid discrete event simulation environment provides flexible and dynamically reconfigurable models of the behavior of components and fluids in the life support systems. Both discrete event and continuous (discrete time) simulation are supported, and flows and pressures are computed globally. This provides fast dynamic simulations of interacting hardware systems in closed loops that can be reconfigured during operations scenarios, producing complex cascading effects of operations and failures. Current object-oriented model libraries support modeling of fluid systems, and models have been developed of physico-chemical and biological subsystems for processing advanced life support gases. In FY01, water recovery system models will be developed.

Malin, Jane T.↗

Analyzing Potential Failures and Effects in a Pilot-Scale Biomass Preprocessing Facility for Improved Reliability

This study demonstrates a failure identification methodology applied to a preprocessing facility generating conversion-ready feedstocks from biomass meeting conversion process critical quality attribute (CQA) specifications. Failure Modes and Effects Analysis (FMEA) was used as an industrially relevant risk analysis approach to evaluate a logging residue preprocessing system to prepare feedstock for pyrolysis conversion. Risk evaluations considered both system-level and operation unit-level assessments considering process efficiency, product quality, cost, sustainability, and safety. Key outputs included estimations of semi-quantitative risk scores for each failure, identification of the failure impacts, identification of failure causes associated with material attributes and process parameters, ranking success rates of failure detection methods, and speculation of potential mitigation strategies for decreasing failure risk scores. Results showed that deviations from moisture specifications had cascading consequences for other CQAs along with process safety implications. Failures linked to fixed carbon specifications carried the highest risk scores for product quality and process efficiency impacts. As increased throughput can be inversely related to meeting product quality specifications; achieving throughput and other material-based CQAs simultaneously will likely require system optimization or prioritization based on system economics. Ultimately, this work successfully demonstrates FMEA as a risk analysis approach for other bioenergy process systems.

09 BIOMASS FUELS↗

Distribution of blackouts in the power grid and the Motter and Lai model

Carreras, Dobson, and colleagues have studied empirical data on the sizes of the blackouts in real grids and modeled them with computer simulations using the direct current approximation. They have found that the resulting blackout sizes are distributed as a power law and suggested that this is because the grids are driven to the self-organized critical state. In contrast, more recent studies found that the distribution of cascades is bimodal resulting in either a very small blackout or a very large blackout, engulfing a finite fraction of the system. Here we reconcile the two approaches and investigate how the distribution of the blackouts changes with model parameters, including the tolerance criteria and the dynamic rules of failure of the overloaded lines during the cascade. Finally, we study the same problem for the Motter and Lai model and find similar results, suggesting that the physical laws of flow on the network are not as important as network topology, overload conditions, and dynamic rules of failure.

24 POWER TRANSMISSION AND DISTRIBUTION↗