Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “limited data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

Dynamic Boundary Microgrids Under Privatization Considerations

Microgrids have physical, electrical, and logical (data, network, and ownership) boundaries. To power unserved customer loads during an outage, microgrids can extend the traditional operational boundaries. This can become complex when considering microgrid-to-microgrid (M2M) interactions where sensitive information such as competitive microgrid operational data is not shared. This work proposes an optimization method coordinated between microgrid controllers and distribution management systems that limits data sharing. The method involves a competitive bidding strategy that maximizes unserved load coverage while minimizing resource utilization and sensitive operational data sharing among entities. The work is validated on a two-microgrid system with photovoltaic and energy storage systems and curves of load derived from real world residential buildings datasets. Results show that the proposed method, when applied for three distinct use cases of energy storage sufficiency to cover the predefined boundary and/or the expanded boundary, can successfully select and bid the available load coverage.

Starke, Michael [ORNL] (ORCID:0000000221211195)↗

Baseline Cost Model for Hydropower: Documentation (2025)

Hydropower currently contributes about 80 GW of conventional and 23 GW of pumped storage capacity to the United States (US) power grid. Previous studies have estimated a considerable amount of remaining US hydropower resources, including non-powered dams (NPD) (Hadjerioua et al., 2012), new stream-reach developments (NSD) (Kao et al., 2014), pumped storage hydropower (PSH) and canal/conduit (Kao et al., 2022). The combined theoretical capacity potential of these various hydropower resources is comparable to the existing US hydropower capacity. There is a continuing interest in developing this hydropower potential, particularly to help meet the increasing demand for electricity. However, available data (Sasthav and Oladosu, 2022) show that the rate of new hydropower development has slowed considerably over time despite the interest of industry stakeholders. This is partly due to the competition from other energy resources and from the highly dispersed nature of remaining hydropower resources, which lead to high information requirements for evaluating the feasibility of potential projects. Cost information provides the most succinct summary of the feasibility of a potential hydropower project required by stakeholders, including developers, investors, policymakers, consumer groups, etc., considering investment options. The best estimates of hydropower costs can be obtained through detailed engineering design and cost assessments of individual projects. However, this approach has high data and resource (time, funds, cross-disciplinary expertise) requirements that render it inapplicable for rapid cost estimation with limited data. Although innovative approaches can overcome some of these impediments (see Oladosu and Ma, 2024 for such an application to potential NPD projects), the development of such approaches still requires significant amounts of resources and are not generally applicable to all hydropower project types. Therefore, statistical and parametric methods using simpler cost specifications remain of significant utility to hydropower stakeholders and are, at the least, complementary to more detailed approaches, particularly when evaluating many potential projects.

13 HYDRO ENERGY↗

Integrated Framework of Multisource Data Fusion for Outage Location in Looped Distribution Systems

Accurate outage location is essential for expediting post-outage power restoration, minimizing outage duration, and enhancing the resilience of distribution networks. With the advent of advanced metering infrastructure, data-driven outage location methods have significantly advanced beyond traditional approaches that rely on manual inspections. However, existing methods still face critical challenges, like reliance on single-source data, limited ability to handle partially observable systems or difficulties with loop networks. To the best of our knowledge, no single approach has comprehensively addressed all of these challenges at once. To this end, this paper proposes a comprehensive multisource data fusion framework for outage locations via probabilistic graph networks. The framework consists of three key phases. First, a novel method for reconstituting distribution networks with loops is developed, transforming looped networks into multiple radial subnetworks that retain all outage causalities of the original network. Second, Bayesian network (BN) models are established for each subnetwork, integrating multiple data sources and network structures. Finally, a joint Gibbs sampling mechanism, featuring forward and backward information flow, is designed to merge data from separate BN models and maximize the utilization of limited evidence, ensuring accurate outage location identification. In conclusion, the framework was validated on two modified public test systems, and comparative studies confirmed its effectiveness.

24 POWER TRANSMISSION AND DISTRIBUTION↗

The Design and Implementation of a Secure Datastore Based on Ethereum Smart Contract

In this paper, we present a secure datastore based on an Ethereum smart contract. Our research is guided by three research questions. First, we will explore to what extend a smart-contract-based datastore should resemble a traditional database system. Second, we will investigate how to store the data in a smart-contract-based datastore for maximum flexibility while minimizing the gas consumption. Third, we seek answers regarding whether or not a smart-contract-based datastore should incorporate complex processing such as data encryption and data analytic algorithms. The proposed smart-contract-based datastore aims to strike a good balance between several constraints: (1) smart contracts are publicly visible, which may create a confidentiality concern for the data stored in the datastore; (2) unlike traditional database systems, the Ethereum smart contract programming language (i.e., Solidity) offers very limited data structures for data management; (3) all operations that mutate the blockchain state would incur financial costs and the developers for smart contracts must make sure sufficient gas is provisioned for every smart contract call, and ideally, the gas consumption should be minimized. Our investigation shows that although it is essential for a smart-contract-based datastore to offer some basic data query functionality, it is impractical to offer query flexibility that resembles that of a traditional database system. Furthermore, we propose that data should be structured as tag-value pairs, where the tag serves as a non-unique key that describes the nature of the value. We also conclude that complex processing should not be allowed in the smart contract due to the financial burden and security concerns. The tag-based secure datastore designed this way also defines its applicative perimeter, i.e., only applications that align with our strategy would find the proposed datastore a good fit. Those that would rather incur higher financial cost for more data query flexibility and/or less user burden on data pre- and post-processing would find the proposed database too restrictive.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Understanding Aitken Mode Aerosol Variability over the Southern Ocean and Antarctica: Insights from Cloud Condensation Nuclei Data

Aitken mode aerosol particles play an important role influencing cloud properties and sustenance, acting as a reservoir of potential cloud condensation nuclei against precipitation scavenging. However, there is limited data on Aitken mode aerosols. In this study, we develop a method to estimate Aitken mode aerosol concentrations and size distribution using cloud condensation nuclei measurements (CCN) and κ-Köhler theory. The performance of this method is evaluated using scanning mobility particle sizer (SMPS) data from recent field campaigns to demonstrate its skills and applicability. The method reasonably estimates Aitken- and accumulation-mode aerosol concentrations, achieving correlations of 0.7–0.9 with only modest biases (mean fractional bias within ±23% for Aitken mode and ±34% for accumulation-mode). This method is further applied to measurements collected over the Southern Ocean and Antarctica in recent years from multiple platforms, including ground sites, aircraft, and ships, to derive Aitken and accumulation-mode aerosol concentrations. Using the derived data, we examine the seasonal cycle, latitudinal variations, and vertical distribution of aerosols. Aitken mode aerosol concentrations are elevated over the Southern Ocean and Antarctica during the austral summer similar to the accumulation mode. In the austral summer, the free troposphere has more Aitken mode aerosols and fewer accumulation mode aerosols than the boundary layer, and thus likely serves as an important source of cloud-forming aerosol while also diluting the accumulation mode.

Kang, Litai [University of Washington] (ORCID:0000↗

Modeling Electric Vehicle Charging Station Siting Suitability with a Focus on Equity

As adoption of electric vehicles increases, the infrastructure to charge them must keep pace. Determining where to add new charging infrastructure is a complex process subject to many factors, including electrical service availability, vehicle dwell time, the type(s) of drivers and vehicles the stations will serve, traffic levels and timing, and land ownership. In addition, advancing social equity is a current priority of federal efforts to invest in electric vehicle charging infrastructure. Conducting Multi-criteria Decision Analysis (MCDA) within Argonne’s Energy Zones Mapping Tool (EZMT) is a useful method for analyzing many of the factors that influence how suitable a location is for potentially adding new charging infrastructure, and we show how equity metrics can be included in the analysis. However, data limitations impose challenges to using MCDA to evaluate and prioritize locations. We use three examples to demonstrate how to use publicly available data and MDCA to analyze different siting objectives. Each example starts with defining a specific objective and ends with how to use the results to identify specific potential locations that could be investigated further. This analysis demonstrates how interested stakeholders can use the EZMT to run the example MCDA models defined in this study, modify them to suit their needs, or create new MCDA models.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Making a Water Data System Responsive to Information Needs of Decision Makers

Evidence-based environmental management requires data that are sufficient, accessible, useful and used. A mismatch between data, data systems, and data needs for decision making can result in inefficient and inequitable capital investments, resource allocations, environmental protection, hazard mitigation, and quality of life. In this paper, we examine the relationship between data and decision making in environmental management, with a focus on water management. We focus on the concept of decision-driven data systems —data systems that incorporate an assessment of decision-makers' data needs into their design. The aim of the research was to examine the process of translating data into effective decision making by engaging stakeholders in the development of a water data system. Using California's legislative mandate for state agencies to integrate existing water and other environmental data as a case study, we developed and applied a participatory approach to inform data-system design and identify unmet data needs. Using workshops and focused stakeholder meetings, we developed 20 diverse use cases to assess data sources, availability, characteristics, gaps, and other attributes of data used for representative decisions. Federal and state agencies made up about 90% of the data sources, and could readily adapt to a federated data system, our recommended model for the state. The remaining 10% of more-specialized data, central to important decisions across multiple use cases, would require additional investment or incentives to achieve data consistency, interoperability, and compatibility with a federated system. Based on this assessment, we propose a typology of different types of data limitations and gaps described by stakeholders. We also propose technical, governance, and stakeholder engagement evaluation criteria to guide planning and building environmental data systems. Data-system governance involving both producers and users of data was seen as essential to achieving workable standards, stable funding, convenient data availability, resilience to institutional change, and long-term buy-in by stakeholders. Our work provides a replicable lesson for using decision-maker and stakeholder engagement to shape the design of an environmental data system, and inform a technical design that addresses both user and producer needs.

Cantor, Alida↗

Efficient Reinforcement Learning for Real-Time Hardware-Based Energy System Experiments: Preprint

In the context of urgent climate challenges and the pressing need for rapid technology development, Reinforcement Learning (RL) stands as a compelling data-driven method for controlling real-world physical systems. However, RL implementation often entails time-consuming and computationally intensive data collection and training processes, rendering them inefficient for real-time applications that lack non-real-time models. To address these limitations, real-time emulation techniques have emerged as valuable tools for the lab-scale rapid prototyping of intricate energy systems. While emulated systems offer a bridge between simulation and reality, they too face constraints, hindering comprehensive characterization, testing, and development. In this research, we construct a surrogate model using limited data from simulated systems, enabling an efficient and effective training process for a Double Deep Q-Network (DDQN) agent for future deployment. Our approach is illustrated through a hydropower application, demonstrating the practical impact of our approach on climate-related technology development.

deep Q-learning↗

Reevaluation of Radiation-Protection Standards for Workers and the Public Based on Current Scientific Evidence

President Trump’s recent executive orders to “Usher in a Nuclear Renaissance,” coupled with the global pledge to triple nuclear energy capacity by 2050, underscore nuclear energy’s importance to national security and economic prosperity. This renewed interest has prompted efforts to spur nuclear-energy deployment, including assessing factors impeding it. One issue that has previously been identified as adding to the cost of nuclear energy is excessively conservative requirements, including those related to radiation protection. This technical review, therefore, examines current radiation-protection standards that were established decades ago when more limited data were available and nuclear-energy expansion was not a national priority. The review focuses on scientific evidence regarding the health effects of ionizing radiation at annual doses of 10,000 mrem or less. The review evaluates epidemiological studies, radiobiological research, and positions of relevant professional organizations to assess whether such doses result in discernable or observable increases in negative health outcomes. The review also surveys the literature related to economic and practical implications of current radiation-protection standards and practices. Based on this assessment, we propose maintaining an annual occupational whole-body dose limit of 5,000 mrem/yr and eliminating all “as low as reasonably achievable” requirements and limits below this threshold. This change could potentially reduce radiation-protection costs by millions of dollars annually for each reactor, as well as decrease the overall costs and correct misconceptions about the risks associated with all nuclear technologies. The evidence further supports future consideration of a 10,000 mrem/yr limit that would maintain appropriate safety margins while further reducing protection costs. Similarly, given the data and that the average annual radiation dose per person in the U.S. is 620 mrem, we believe the public dose limits of 100 mrem/yr are unnecessarily restrictive; increasing to 500 mrem/yr would maintain substantial safety margins—a factor of 10 below the occupational limit—while reducing regulatory burdens and associated bureaucracy. While we acknowledge ongoing scientific debate and encourage continued research on the health effects of ionizing radiation, our review indicates current frameworks are overly conservative. These overly stringent limits not only impose unnecessary economic burdens without corresponding health benefits but also divert safety focus and resources from more important considerations. Although this study was motivated by nuclear-power considerations, reforms to radiation-protection requirements have significant positive implications for other areas, such as nuclear medical applications, environmental remediation, nuclear-waste management and disposal, and industrial applications of nuclear technologies.

61 RADIATION PROTECTION AND DOSIMETRY↗

Correlated Trajectory Uncertainty for Adaptive Sequential Decision Making

One of the great challenges with decision making tasks on real world systems is the fact that data is sparse and acquiring additional data is expensive. In these cases, it is often crucial to make a model of the environment to assist in making decisions. At the same time, limited data means that learned models are erroneous, making it just as important to equip the model with good predictive uncertainties. In the context of learning sequential decision making policies, these uncertainties can prove useful for informing which data to collect for the greatest improvement in policy performance \citep{mehta2021experimental, mehta2022exploration} or informing the policy about unsure regions of state and action space to avoid during test time \citep{yu2020mopo}. Additionally, assuming that realistic samples of the environment can be drawn, an adaptable policy can be trained that attempts to make optimal decisions for any given possible instance of the environment \citep{ghosh2022offline, chen2021offline}. In this work, we examine the so-called ``probabilistic neural network'' (PNN) model that is ubiquitous in model-based reinforcement learning (MBRL) works. We argue that while PNN models may have good marginal uncertainties, they form a distribution of non-smooth transition functions. Not only are these samples unrealistic and may hamper adaptability, but we also assert that this leads to poor uncertainty estimates when predicting multiple step trajectory estimates. To address this issue, we propose a simple sampling method that can be implemented on top of pre-existing models.We evaluate our sampling technique on a number of environments, including a realistic nuclear fusion task, and find that, not only do smooth transition function samples produce more calibrated uncertainties, but they also lead to better downstream performance for an adaptive policy.

Offline Reinforcement Learning↗

AGR-5/6/7 Experiment Preliminary Findings

AGR-5/6/7 in-pile data have been published: Fuel irradiation conditions, Fission gas release-rate-to-birth-rate (R/B) ratios, PIE and safety testing are still in progress and only limited data have been published, and preliminary results.

11 - NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Micrometer: Micromechanics transformer for predicting full field mechanical responses of heterogeneous materials

Predicting mechanical responses of heterogeneous materials across scales remains a significant challenge. Traditional computational methods often struggle with complex and multiscale nature of these materials, limiting their effectiveness in real-world applications. Here, in this paper, we introduce Micrometer, a vision transformer based deep learning model designed to predict full field mechanical responses of heterogeneous materials, bridging the gap between computer vision and solid mechanics problems. We show that Micrometer, trained on a large-scale high-resolution dataset of 2D fiber-reinforced composites, can achieve state-of-the-art performance in predicting microscale strain fields across a wide range of material properties and loading conditions. Our model demonstrates accuracy and computational efficiency in applications such as computational homogenization and multiscale modeling, reducing computational time by up to two orders of magnitude compared to conventional numerical solvers while maintaining less than 1 % errors in predicting macroscale stress fields. Furthermore, we showcase Micrometer’s adaptability through transfer learning experiments on new materials with limited data, highlighting its potential to tackle diverse scenarios in computational solid mechanics. These results represent a significant step towards AI-driven innovation in materials science, addressing the limitations of traditional numerical methods and paving the way for more efficient simulations of heterogeneous materials across various industrial applications.

Composite materials↗

The expressivity of classical and quantum neural networks on entanglement entropy

Abstract Analytically continuing the von Neumann entropy from Rényi entropies is a challenging task in quantum field theory. While then-th Rényi entropy can be computed using the replica method in the path integral representation of quantum field theory, the analytic continuation can only be achieved for some simple systems on a case-by-case basis. In this work, we propose a general framework to tackle this problem using classical and quantum neural networks with supervised learning. We begin by studying several examples with known von Neumann entropy, where the input data is generated by representing$${\text {Tr}}\rho _A^n$$ Tr ρ A n with a generating function. We adopt KerasTuner to determine the optimal network architecture and hyperparameters with limited data. In addition, we frame a similar problem in terms of quantum machine learning models, where the expressivity of the quantum models for the entanglement entropy as a partial Fourier series is established. Our proposed methods can accurately predict the von Neumann and Rényi entropies numerically, highlighting the potential of deep learning techniques for solving problems in quantum information theory.

Physics↗

Commercial, industrial, and institutional discount rate estimation for efficiency standards analysis: Sector-level data 1998–2021

Underlying each of the Department of Energy’s (DOE’s) federal appliance and equipment energy conservation standards are a set of complex analyses of the projected costs and benefits of regulation. Any new or amended standard must be designed to achieve significant additional energy conservation, provided that it is technologically feasible and economically justified (42 U.S.C. 6295(o)(2)(A)). DOE determines economic justification based on whether the benefits exceed the burdens, considering a variety of factors, including the economic impact of the standard on consumers of the product and the savings in lifetime operating cost compared to any increase in price or maintenance expenses (42 U.S.C. 6295(o)(2)(B)). As part of this determination, DOE conducts a Life-Cycle Cost (LCC) analysis, which models the combined impact of appliance first cost and operating cost changes on a representative commercial building sample in order to identify the fraction of customers achieving LCC savings or incurring net cost at the considered efficiency levels. Thus, the commercial discount rate value(s) used to calculate the present value of energy cost savings within the LCC model implicitly plays a role in estimating the economic impact of potential standard levels. This report provides an in-depth discussion of the commercial discount rate estimation process. It is an update to previous reports on estimating commercial discount rates from firm-level financial data (Fujita, 2016). Major topics covered in this report include: Discount rate estimation methods and rationale; -Data sources used and data limitations; -Discount rate distributions for use in standards analysis; -Discount rate estimation methods and distributions specific to the small business subgroup analysis. Going forward, this report will be updated as data allow and analyses necessitate.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Domain-aware Control-oriented Neural Models for Autonomous Underwater Vehicles

Conventional physics-based modeling is a time-consuming bottleneck in control design for complex nonlinear systems like autonomous underwater vehicles (AUVs). In contrast, purely data-driven models, require a large number of observations and lack operational guarantees for safety-critical systems. Data-driven models leveraging available partially characterized dynamics have potential to provide reliable systems models in a typical data-limited scenario for high value complex systems, thereby avoiding months of expensive expert modeling time. In this work we explore this middle-ground between expert-modeled and pure data-driven modeling. We present control-oriented parametric models with varying levels of domain-awareness that exploit known system structure and prior physics knowledge to create constrained deep neural dynamical system models. We employ universal differential equations to construct data-driven blackbox and graybox representations of the AUV dynamics. In addition, we explore a hybrid formulation that explicitly models the residual error related to imperfect graybox models. We compare the prediction performance of the learned models for different distributions of initial conditions and control inputs to assess their suitability for control.

Shaw Cortez, Wenceslao E.↗

Mixture Model for Refrigerant Pairs R-32/1234yf, R-32/1234ze(E), R-1234ze(E)/227ea, R-1234yf/152a, and R-125/1234yf

In this work, thermodynamic models based on the corresponding states framework with departure terms are developed for the refrigerant pairs R-32/1234yf, R-32/1234ze(E), R-1234ze(E)/227ea, R-1234yf/152a, and R-125/1234yf. These models are based on new measurements of density, speed of sound, and phase equilibria, combined with the data available in the literature. The model for R-32/1234yf is most comprehensive in its data coverage, with speed of sound deviations within 1%, density deviations within 0.1%, and bubble- and dew-point pressure deviations within 1%. In conclusion, the other mixtures have generally more limited data availability but a similar goodness of fit.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

A VOI Web Application for Distinct Geothermal Domains: Statistical Evaluation of Different Data Types within the Great Basin

The Great Basin region contains different domains that have different structural and hydrothermal flow patterns. Depending on the characteristics of these patterns, certain data types may be more successful at detecting hidden geothermal resources. In this paper, we quantitatively evaluate if certain data types are more successful in certain domains. Given different aquifer, strain and structural conditions, we explore which data types statistically reveal positively labeled geothermal sites. We utilize value of information (VOI) metrics to help quantify the reliability of data types to discriminate against "positive" and "negative" labeled geothermal sites. We also evaluate how kernel density estimation can help generalize the statistics that inform VOI, which is necessary given the limited data in geothermal exploration. Except for the Carbonate Aquifer, the highest ranking of the Vimperfect is the Local Structural Setting. Next, the slip and dilation tendency is first for Carbonate Aquifer and second for Central Nevada Seismic Belt and Western Great Basin. For the Carbonate Aquifer, heat flow is has the lowest Vimperfect value compared to the other three domains, which is consistent with the understanding of how heat flow measurements are masked by regional groundwater flow.

Bayesian analysis↗

Journey over Destination: Dynamic Sensor Placement Enhances Generalization

Reconstructing complex, high-dimensional global fields from limited data points is a challenge across various scientific and industrial domains. This is particularly important for recovering spatio-temporal fields using sensor data from, for example, laboratory-based scientific experiments, weather forecasting, or drone surveys. Given the prohibitive costs of specialized sensors and the inaccessibility of
certain regions of the domain, achieving full field coverage is typically not feasible. Therefore, the development of machine learning algorithms trained to reconstruct fields given a limited dataset is of critical importance. In this study, we introduce a general
approach that employs moving sensors to enhance data exploitation during the training of an attention based neural network, thereby improving field reconstruction. The training of sensor locations is accomplished using an end-to-end workflow, ensuring
differentiability in the interpolation of field values associated to the sensors, and is simple to implement using differentiable programming. Additionally, we have incorporated a correction mechanism to prevent sensors from entering invalid regions within the domain. We evaluated our method using two distinct datasets; the results show that our approach enhances learning, as evidenced by improved test scores.

54 ENVIRONMENTAL SCIENCES↗