Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Tolerance Bounds”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 181 records · Page 10

Uncertain quantum computing futures and potential energy and physical resource impacts at scale

Considerable attention has recently focused on the vast energy and water demands of supercomputing, namely large-scale data centers that underpin artificial intelligence (AI), one of the great disruptors of contemporary society. Looking ahead some years from now, quantum computing is poised to disrupt established computing paradigms once again. Scientists and engineers are now working intensely to bring this century-old dream of physicists to fruition. Yet, as quantum computers begin to be integrated with classical supercomputing architectures, the implications for energy and physical resource use also need to be understood, especially how they compare to today’s AI data centers. These impacts have not yet been quantified by the research community – a notable gap in the literature, even if commercial-scale deployment of Quantum-Accelerated Computing Infrastructure (QuACI) is not expected for a few more years. This study is the first to conduct such an assessment. Using publicly available information from academic sources and private industry, we characterize multiple configurations of superconducting qubit-based, fault-tolerant quantum computers (FTQC) that could plausibly be deployed at scale in the 2030s and into the 2040s. By parameterizing these FTQC systems at a process level, we conduct a prospective scenario analysis to quantify their energy and physical resource needs. While these estimates are uncertain, given the current state of quantum technologies and their unknown future trajectories, important insights can already be drawn. One key finding is that while the electricity needs for a fleet of FTQCs are within the bounds of previous modeling studies that have explored high electricity demand futures, the needs for certain physical resources, namely water and helium-3, could pose bottlenecks to QuACI scale-up.

Computing↗

Improved ion heating in fast ignition by pulse shaping

The fast ignition paradigm for inertial fusion offers increased gain and tolerance of asymmetry by compressing fuel at low entropy and then quickly igniting a small region. Because this hotspot rapidly disassembles, the ions must be heated to ignition temperature as quickly as possible, but most ignitor designs directly heat electrons. A constant-power ignitor pulse, which is generally assumed, is suboptimal for coupling energy from electrons to ions. Using a simple model of a hotspot in isochoric plasma, a novel pulse shape to maximize ion heating is presented in analytical form. Bounds are derived on the maximum ion temperature attainable by electron heating only. Moreover, arranging for faster ion heating allows a smaller hotspot, improving fusion gain. As a result, under representative conditions, the optimized pulse can reduce ignition energy by over 20%.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Prediction of Performance Variation Caused by Manufacturing Tolerances and Defects in Gas Diffusion Electrodes of Phosphoric Acid (PA)–Doped Polybenzimidazole (PBI)-Based High-Temperature Proton Exchange Membrane Fuel Cells

The automated process of coating catalyst layers on gas diffusion electrodes (GDEs) for high-temperature proton exchange membrane fuel cells results inherently into a number of defects. These defects consist of agglomerates in which the platinum sites cannot be accessed by phosphoric acid and which are the consequence of an inconsistent coating, uncoated regions, scratches, knots, blemishes, folds, or attached fine particles—all ranging from μm to mm size. These electrochemically inactive spots cause a reduction of the effective catalyst area per unit volume (cm2/cm3) and determine a drop in fuel cell performance. A computational fluid dynamics (CFD) model is presented that predicts performance variation caused by manufacturing tolerances and defects of the GDE and which enables the creation of a six-sigma product specification for Advent phosphoric acid (PA)-doped polybenzimidazole (PBI)-based membrane electrode assemblies (MEAs). The model was used to predict the total volume of defects that would cause a 10% drop in performance. It was found that a 10% performance drop at the nominal operating regime would be caused by uniformly distributed defects totaling 39% of the catalyst layer volume (~0.5 defects/μm2). The study provides an upper bound for the estimation of the impact of the defect location on performance drop. It was found that the impact on the local current density is higher when the defect is located closer to the interface with the membrane. The local current density decays less than 2% in the presence of an isolated defect, regardless of its location along the active area of the catalyst layer.

Gurau, Vladimir (ORCID:0000000327429061)↗

A Confidence-Based Approach to Including Survivors in a Probabilistic TID Failure Assessment

A probabilistic total ionizing dose (TID) failure assessment is extended to include survivor data, enabling the bounding of failure probability to a desired confidence level (CL) without failure data. The extension provides an avenue for analyzing microelectronics tested for TID without reaching a failure mode, a scenario often encountered by missions utilizing commercial-off-the-shelf (COTS) technologies. Using the type-I censored likelihood formulation and a realistic upper bound on expected device performance, the failure probability space is bounded by confidence contours within the context of a variable environment. The framework accommodates any type of distribution assumed for the part failure or the environment under consideration. Furthermore, the framework can be utilized pre-emptively to plan future device TID tests, minimizing costs while meeting survival requirements. Heritage data may also be used as survivors to further minimize testing costs when parts are from the same lot, but the amount of constraint derived from heritage is limited. Altogether, the framework enables a formal, mathematically rigorous analysis of radiation tolerant devices tested to a maximum dose, as well as flight heritage, in a hardness assurance methodology.

confidence↗

On-board closed-loop congestion control for satellite based packet switching networks

NASA LeRC is currently investigating a satellite architecture that incorporates on-board packet switching capability. Because of the statistical nature of packet switching, arrival traffic may fluctuate and thus it is necessary to integrate congestion control mechanism as part of the on-board processing unit. This study focuses on the closed-loop reactive control. We investigate the impact of the long propagation delay on the performance and propose a scheme to overcome the problem. The scheme uses a global feedback signal to regulate the packet arrival rate of ground stations. In this scheme, the satellite continuously broadcasts the status of its output buffer and the ground stations respond by selectively discarding packets or by tagging the excessive packets as low-priority. The two schemes are evaluated by theoretical queuing analysis and simulation. The former is used to analyze the simplified model and to determine the basic trends and bounds, and the later is used to assess the performance of a more realistic system and to evaluate the effectiveness of more sophisticated control schemes. The results show that the long propagation delay makes the closed-loop congestion control less responsive. The broadcasted information can only be used to extract statistical information. The discarding scheme needs carefully-chosen status information and reduction function, and normally requires a significant amount of ground discarding to reduce the on-board packet loss probability. The tagging scheme is more effective since it tolerates more uncertainties and allows a larger margin of error in status information. It can protect the high-priority packets from excessive loss and fully utilize the downlink bandwidth at the same time.

Chu, Pong P.↗

Robust control design for aerospace applications

Time-domain control design for stability robustness of linear systems with structured uncertainty is addressed. Upper bounds on the linear perturbation of an asymptotically stable linear system are obtained, making it possible to maintain stability by using the structural information of the uncertainty. A quantitative measure called the stability robustness index is introduced and used to design controllers for robust stability. The proposed state feedback control design algorithm can be used, for a given set of perturbations, to select the range of control effort for which the system is stability-robust. Conversely it can be used, for a given control effort, to determine the size of the tolerable perturbation. The algorithm is illustrated with examples from aircraft control and large-space-structure control problems.

Yedavalli, Rama K.↗

Enhancing Distribution System Resilience: A First-Order Meta-RL Algorithm for Critical Load Restoration

The increasing frequency of extreme events and the integration of distributed energy resources (DERs) into modern grids have elevated the need for resilient and efficient critical load restoration strategies in distribution systems. However, the stochastic nature of renewable DERs, limited energy resource availability and the intricate nonlinearities inherent in complex grid control problem make the problem challenging. Although reinforcement learning (RL) and warm-start RL methods have shown promising results, their performance often falls short in rapidly adapting to new, unseen situations and typically requires exhaustive problem-specific tuning. To address these gaps, we propose a First-Order Meta-based RL (FOM-RL) algorithm within an online framework for adaptive and robust critical load restoration. By harnessing local DERs as the enabling technology, FOM-RL allows the RL agent to swiftly adapt to new unseen scenarios by leveraging previously acquired knowledge of different tasks. Experimental results provide evidence that proposed algorithm learns more efficiently and showcases generalization capabilities across diverse set of operational scenarios. Moreover, a rigorous theoretical analysis yields a tight sublinear regret bound, sensitive to temporal variability, with a task-averaged optimality gap bounded by O(VM+D*/(Tsquare root(M))). These results suggest that optimality improves with task similarity and an increased number of tasks M, reaffirming the efficacy and scalability of the proposed approach in addressing the complexities of critical load restoration in distribution systems.

complexity theory↗

Investigating the Role of Accident Tolerant Cladding on Source Term Reduction for High-Burnup PWRs Using MELCOR

The use of accident tolerant fuel (ATF) cladding can increase coping times during and beyond design basis accidents. While such gains may be incremental, they provide a margin that can potentially be recovered to enable high-burnup (HBU) operation. Realizing such a margin requires demonstrating that the combination of HBU and ATF has not led to an overall increase in source term. This study investigates the influence of cladding technology (Zr-based, Cr-coated Zr, and FeCrAl) and fuel cycle length (18 and 24 months) on radiological dose at the boundary of the exclusion zone for a four-loop pressurized water reactor to investigate whether ATF claddings can provide such benefits. We analyze a recovered large break loss-of-coolant accident scenario to investigate the impact of transient timescale on the benefits of such coping time increases. The simulations have been performed using the MELCOR and MELCOR Accident Consequence Code System codes. For the cases analyzed, increased fuel cycle length did not necessarily increase radionuclide release and hydrogen generation, as these were found to be sensitive to the core power distribution. Similarly, off-site dose consequence is dominated by short-lived radionuclides that tend to saturate earlier in the burnup, so higher burnup operation did not necessarily increase the source term for the phenomena and transients analyzed here. Delays in recovery of the lowpressure safety injection system increase hydrogen production and radionuclide release, especially between 780 s and 1620 s, due to the nonlinear oxidation and core degradation behavior. Results show that Cr-coated Zr enhances safety by delaying heatup and gap release. Here, when uncertainty propagation on oxidation properties is considered, FeCrAl exhibits the lowest overall radionuclide release and off-site dose throughout the spectrum. However, while the considered “base model” performance is superior under delayed injection scenarios, upper-bound cases display hydrogen generation risk comparable to the Zr-based cladding.

Accident Tolerant Fuel↗

Flight critical system design guidelines and validation methods

Efforts being expended at NASA-Langley to define a validation methodology, techniques for comparing advanced systems concepts, and design guidelines for characterizing fault tolerant digital avionics are described with an emphasis on the capabilities of AIRLAB, an environmentally controlled laboratory. AIRLAB has VAX 11/750 and 11/780 computers with an aggregate of 22 Mb memory and over 650 Mb storage, interconnected at 256 kbaud. An additional computer is programmed to emulate digital devices. Ongoing work is easily accessed at user stations by either chronological or key word indexing. The CARE III program aids in analyzing the capabilities of test systems to recover from faults. An additional code, the semi-Markov unreliability program (SURE) generates upper and lower reliability bounds. The AIRLAB facility is mainly dedicated to research on designs of digital flight-critical systems which must have acceptable reliability before incorporation into aircraft control systems. The digital systems would be too costly to submit to a full battery of flight tests and must be initially examined with the AIRLAB simulation capabilities.

Holt, H. M.↗

Fault-tolerant resource comparison of qudit and qubit encodings for diagonal quadratic operators

Finite local Hilbert-space truncations arise naturally in quantum simulations of lattice field theories and motivate qudit encodings, but their fault-tolerant advantage over qubit encodings remains unclear. We compare the non-Clifford cost of implementing quadratic diagonal evolutions, exemplified by 𝑈 = 𝑒$^{−𝑖⁢𝑡⁢𝜙^2_𝑥}$ in a uniform field-amplitude discretization of a real scalar field, using either one logical 𝑑-level qudit or 𝑛 𝑏 = ⌈log 2⁡ 𝑑⌉ logical qubits. We analyze two standard settings: product-formula simulation and linear combination of unitaries (LCU) per block encoding, taking the resource metric to be the number of non-Clifford gates after synthesis into a discrete logical gate set. Because tight synthesis bounds for general single-qudit rotations are not known, we express the qudit constructions in terms of embedded two-level SU⁡(2) rotations and derive explicit finite-𝑑 break-even conditions for their synthesis cost; these serve as compiler targets for when qudit encodings can outperform the qubit baseline. Within the constructive models studied here, product-formula implementations would require an exponentially stronger per-primitive synthesis advantage for qudits to win asymptotically, while in the LCU setting the qubit encoding is asymptotically cheaper in 𝑑. Nevertheless, the finite-𝑑 threshold analysis identifies low-dimensional regions in which qudits can yield meaningful constant-factor savings, particularly for LCU-based implementations. As a secondary analysis of the LCU construction, we use an idealized negligible-overhead qubit-qudit code-switching model to give an absolute 𝑇-count comparison and reinterpret the savings as an allowable per-switch overhead budget.

Godwood, Samuel [Univ. of Liverpool (United Kingdo↗

A Self-Stabilizing Hybrid-Fault Tolerant Synchronization Protocol

In this report we present a strategy for solving the Byzantine general problem for self-stabilizing a fully connected network from an arbitrary state and in the presence of any number of faults with various severities including any number of arbitrary (Byzantine) faulty nodes. Our solution applies to realizable systems, while allowing for differences in the network elements, provided that the number of arbitrary faults is not more than a third of the network size. The only constraint on the behavior of a node is that the interactions with other nodes are restricted to defined links and interfaces. Our solution does not rely on assumptions about the initial state of the system and no central clock nor centrally generated signal, pulse, or message is used. Nodes are anonymous, i.e., they do not have unique identities. We also present a mechanical verification of a proposed protocol. A bounded model of the protocol is verified using the Symbolic Model Verifier (SMV). The model checking effort is focused on verifying correctness of the bounded model of the protocol as well as confirming claims of determinism and linear convergence with respect to the self-stabilization period. We believe that our proposed solution solves the general case of the clock synchronization problem.

Malekpour, Mahyar R.↗

Evaluation of the Sparton tight-tolerance AXBT

Forty-six near-simultaneous pairs of conductivity - temperature - depth (CTD) and Sparton 'tight tolerance' air expendable bathythermograph (AXBT) temperature profiles were obtained in summer 1991 from a location in the Sargasso Sea. The data were analyzed to assess the temperature and depth accuracies of the Sparton AXBTs. The tight-tolerance criterion was not achieved using the manufacturer's equations but may have been achieved using customized equations computed from the CTD data. The temperature data from the customized equations had a one standard deviation error of 0.13 C. A customized elapsed fall time-to-depth conversion equation was found to be z = 1.620t - 2.2384 x 10(exp -4) t(exp 2) + 1.291 x 10(exp -7) t(exp 3), with z the depth in meters and t the elapsed fall time after probe release in seconds. The standard deviation of the depth error was about 5 m; a rule of thumb for estimating maximum bounds on the depth error below 100 m could be expressed as +/-2% of depth or +/- 10 m, whichever is greater. This equation gave greater depth accuracy than either the manufacturer's supplied equation or the navy standard equation.

Boyd, Janice D.↗

FPDetect: Efficient Reasoning About Stencil Programs Using Selective Direct Evaluation

We present FPDetect, a low-overhead approach for detecting logical errors and soft errors affecting stencil computations without generating false positives. We develop an offline analysis that tightly estimates the number of floating-point bits preserved across stencil applications. This estimate rigorously bounds the values expected in the data space of the computation. Violations of this bound can be attributed with certainty to errors. FPDetect helps synthesize error detectors customized for user-specified levels of accuracy and coverage. FPDetect also enables overhead reduction techniques based on deploying these detectors coarsely in space and time. Experimental evaluations demonstrate the practicality of our approach.

97 MATHEMATICS AND COMPUTING↗

Topography and functional traits shape the distribution of key shrub plant functional types in low-Arctic tundra

The expansion of shrubs in the Arctic tundra fundamentally modifies land-atmosphere interactions. However, it remains unclear how shrub distribution and expansion differ across key species due to challenges with discriminating tundra plant species at regional scales. Here, we combined multi-scale, multi-platform remote sensing and in situ trait measurements to elucidate the distribution patterns and primary controls of two representative deciduous-tall-shrub (DTS) genera, Alnus and Salix, in low-Arctic tundra. We show that topographic features were a key control on DTSs, creating heterogeneous, but predictable distributions of Alnus and Salix fractional cover (fCover). Alnus was more tolerant of elevation and slope and was found on hilly uplands (slope >10°) within a specific elevational band (200–400 m above sea level [MSL]). In contrast, Salix occurred at lower elevations (50–300 m MSL) on gentler slopes (3-10°) and required adequate soil moisture associated with its profligate water use. We also show that niche differentiation between Alnus and Salix changed with patch size, where larger patches were more specialized in resource requirements than individual plants of Alnus and Salix. To understand what constrains the growth of DTSs at locations with low fCover, we developed environmental limiting factor models, which showed that topography limits the upper bound of Alnus and Salix fCover in 69.2% and 48.7% of the landscape, respectively. These findings highlight a critical need to better understand and represent topography-controlled processes and functional traits in regulating shrub distribution, as well as a need for more detailed species classification to predict shrubification in the Arctic.

alder↗

PROTEUS: Machine Learning Driven Resilience for Extreme-scale Systems

The objective of this project is to design, develop, and evaluate scalable software to enhance resilience, data checkpointing, program restart, and analysis. The proposed tasks are to 1) develop scalable machine learning techniques to learn temporal change patterns in a scalable and in-situ manner, and to minimize data movement and maximize learning locally closest to data; 2) design a concise data representation and indexing mechanism to capture the distribution of changes in data that can guarantee point-wise user-defined tolerable errors while reducing the data storage requirements by an order of magnitude or more; 3) develop data reduction techniques as library modules; 4) exploit local SSD for minimizing data movement in storage hierarchy; 5) develop anomaly detection algorithms that can predict corruptions based on learning of emerging patterns; 6) develop software libraries to be incorporated within widely used data formats and APIs; and 7) evaluate the proposed software using DOE scientific applications. The outcomes of the proposed work are to satisfy many synergistic data reduction and resilience requirements for large-scale data intensive applications executed on extreme-scale computing systems. The developed mechanism for error-bound data approximation is directly applicable to existing scientific applications. Through machine learning from historical events and change distribution, this work will enable anomaly detection for DOE computer facility.

97 MATHEMATICS AND COMPUTING↗

Calculation of Nuclear Reactor Cooling Tower Performance With Limited Data Streams

Monitoring of cooling tower performance in a nuclear reactor facility is necessary to ensure safe operation; however, instrumentation for measuring performance characteristics can be difficult to install and may malfunction or break down over long duration experiments. This paper describes employing a thermodynamic approach to quantify cooling tower performance, the Merkel model, which requires only five parameters, namely, inlet water temperature, outlet water temperature, liquid mass flowrate, gas mass flowrate, and wet bulb temperature. Using this model, a general method to determine cooling tower operation for a nuclear reactor was developed in situations when neither the outlet water temperature nor gas mass flowrate are available, the former being a critical piece of information to bound the Merkel integral. Furthermore, when multiple cooling tower cells are used in parallel (as would be in the case of large-scale cooling operations), only the average outlet temperature of the cooling system is used as feedback for fan speed control, increasing the difficulty of obtaining the outlet water temperature for each cell. To address these shortcomings, this paper describes a method to obtain individual cell outlet water temperatures for mechanical forced-air cooling towers via parametric analysis and optimization. In this method, the outlet water temperature for an individual cooling tower cell is acquired as a function of the liquid-to-gas ratio (L/G). Leveraging the tight tolerance on the average outlet water temperature, an error function is generated to describe the deviation of the parameterized L/G to the highly controlled average outlet temperature. The method was able to determine the gas flowrate at rated conditions to be within 3.9% from that obtained from the manufacturer’s specification, while the average error for the four individual cooling cell outlet water temperatures were 1.6 °C, -0.5 °C, -1.0 °C, and 0.3 °C.

21 SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLANTS↗

Deciphering Microbial Metal Toxicity Responses using RB-TnSeq and Activity-Based Metabolomics

To uncover metal toxicity targets and defense mechanisms of the facultative anaerobe Pantoea sp. strain MT58 (MT58), we used a multiomic strategy combining two global techniques, random bar code transposon site sequencing (RB-TnSeq) and activity-based metabolomics. MT58 is a metal-tolerant Oak Ridge Reservation (ORR) environmental isolate that was enriched in the presence of metals at concentrations measured in contaminated groundwater at an ORR nuclear waste site. The effects of three chemically different metals found at elevated concentrations in the ORR contaminated environment were investigated: the cation Al 3+ , the oxyanion CrO 4 2- , and the oxycation UO 2 2+ . Both global techniques were applied using all three metals under both aerobic and anaerobic conditions to elucidate metal interactions mediated through the activity of metabolites and key genes/proteins. These revealed that Al 3+ binds intracellular arginine, CrO 4 2- enters the cell through sulfate transporters and oxidizes intracellular reduced thiols, and membrane-bound lipopolysaccharides protect the cell from UO 2 2+ toxicity. In addition, the Tol outer membrane system contributed to the protection of cellular integrity from the toxic effects of all three metals. Likewise, we found evidence of regulation of lipid content in membranes under metal stress. Individually, RB-TnSeq and metabolomics are powerful tools to explore the impact various stresses have on biological systems. In this work, we show that together they can be used synergistically to identify the molecular actors and mechanisms of these pertubations to an organism, furthering our understanding of how living systems interact with their environment.

59 BASIC BIOLOGICAL SCIENCES↗

A Self-Stabilizing Hybrid Fault-Tolerant Synchronization Protocol

This paper presents a strategy for solving the Byzantine general problem for self-stabilizing a fully connected network from an arbitrary state and in the presence of any number of faults with various severities including any number of arbitrary (Byzantine) faulty nodes. The strategy consists of two parts: first, converting Byzantine faults into symmetric faults, and second, using a proven symmetric-fault tolerant algorithm to solve the general case of the problem. A protocol (algorithm) is also present that tolerates symmetric faults, provided that there are more good nodes than faulty ones. The solution applies to realizable systems, while allowing for differences in the network elements, provided that the number of arbitrary faults is not more than a third of the network size. The only constraint on the behavior of a node is that the interactions with other nodes are restricted to defined links and interfaces. The solution does not rely on assumptions about the initial state of the system and no central clock nor centrally generated signal, pulse, or message is used. Nodes are anonymous, i.e., they do not have unique identities. A mechanical verification of a proposed protocol is also present. A bounded model of the protocol is verified using the Symbolic Model Verifier (SMV). The model checking effort is focused on verifying correctness of the bounded model of the protocol as well as confirming claims of determinism and linear convergence with respect to the self-stabilization period.

Malekpour, Mahyar R.↗