Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “limited data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

High-throughput spin-bath characterization of spin defects in semiconductors

Detailed knowledge of the local environments of spin defects in semiconductors, such as nitrogenvacancy (NV) centers in diamond or divacancies in silicon carbide, is crucial for optimizing control and entanglement protocols in quantum sensing and information applications. However, at present a direct experimental characterization of individual defect environments is not scalable, as conventional spin-bath measurements are time consuming and difficult to automate. Achieving high-throughput characterization requires short experiments to probe the spin bath. However, with fewer and noisier measurements, the inverse problem of recovering spin-bath properties from measured data becomes ill posed, with multiple spin baths having a high likelihood of yielding the same data. In this work, we present a set of computational tools to resolve the ill-posed inverse problem of recovering the atomic positions and hyperfine couplings of random nuclei surrounding spin defects from sparse, noisy experimental coherence data, which can be obtained in hours. Here, we use a trans-dimensional Bayesian approach that incorporates ab initio data to yield full posterior distributions over nuclear spin environments, enabling robust recovery from limited data. We also provide practical tools and guidelines to determine the limits of detectability for hyperfine couplings under specific dynamical decoupling sequences and sampling conditions. In addition, we demonstrate how the tools developed here, in combination with ab initio simulations of spin baths, can guide the design of efficient experimental protocols for application-specific high-throughput screening. To showcase the utility of our approach, we apply it to design fast dynamical decoupling experiments to characterize the spin baths often individual NV centers in diamond. While the primary focus is on accelerating spin-bath characterization of spin defects, this Bayesian approach also lays the foundation for digital-twin studies of spin defects, where a virtual model of the spin-defect system evolves in real time with ongoing experimental measurements. Together, the set of tools we designed and applied paves the way for scalable deployment of spin defects in semiconductors for quantum sensing and information applications.

Bayesian methods↗

To Fail or not to Fail: An Exploration of Machine Learning Techniques for Predictive Maintenance

Predictive maintenance refers to the ability to predict when machinery or systems need to be maintained. Making an accurate prediction is quite challenging given the costs for both over-estimating (unnecessary maintenance and reduction in availability of assets) and under-estimating (untimely breakdowns and possible loss of equipment or lives). To address these challenges researchers were able to develop new approaches for analyzing oil samples taken extracting samples from oil-wetted machinery that may provide information critical to developing predictive capabilities. We consider the problem from both supervised (though data limited) and unsupervised approaches and provide a first look into a data driven approach for identification of condition indicators. Through this work we identify a collection of candidate features that can form the basis of condition indicators for both a high level discrimination of failure vs. normal operation as well as a set for potential failure mode identification. Finally, we present an anomaly detection framework for detecting failures which can be a viable solution for an onboard analysis tool in deployed systems.

predictive maintenance, anomaly detection, Laserne↗

DOME: Directional medical embedding vectors from Electronic Health Records

Motivation: The increasing availability of Electronic Health Record (EHR) systems has created enormous potential for translational research. Recent developments in representation learning techniques have led to effective large-scale representations of EHR concepts along with knowledge graphs that empower downstream EHR studies. However, most existing methods require training with patient-level data, limiting their abilities to expand the training with multi-institutional EHR data. On the other hand, scalable approaches that only require summary-level data do not incorporate temporal dependencies between concepts. Methods: We introduce a DirectiOnal Medical Embedding (DOME) algorithm to encode temporally directional relationships between medical concepts, using summary-level EHR data. Specifically, DOME first aggregates patient-level EHR data into an asymmetric co-occurrence matrix. Then it computes two Positive Pointwise Mutual Information (PPMI) matrices to correspondingly encode the pairwise prior and posterior dependencies between medical concepts. Following that, a joint matrix factorization is performed on the two PPMI matrices, which results in three vectors for each concept: a semantic embedding and two directional context embeddings. They collectively provide a comprehensive depiction of the temporal relationship between EHR concepts. Results: We highlight the advantages and translational potential of DOME through three sets of validation studies. First, DOME consistently improves existing direction-agnostic embedding vectors for disease risk prediction in several diseases, for example achieving a relative gain of 5.5% in the area under the receiver operating characteristic (AUROC) for lung cancer. Second, DOME excels in directional drug-disease relationship inference by successfully differentiating between drug side effects and indications, correspondingly achieving relative AUROC gain over the state-of-the-art methods by 10.8% and 6.6%. Finally, DOME effectively constructs directional knowledge graphs, which distinguish disease risk factors from comorbidities, thereby revealing disease progression trajectories. The source codes are provided at https://github.com/celehs/Directional-EHRembedding.

60 APPLIED LIFE SCIENCES↗

Database of low‐temperature absorption and fluorescence spectra of native photosynthetic tetrapyrrole macrocycles

Low-temperature (77 K) absorption and fluorescence spectra of 12 naturally occurring photosynthetic tetrapyrrole macrocycles have been recorded in a frozen glass (2-methyltetrahydrofuran). The compounds encompass distinct chromophore classes: porphyrin, chlorophyll c 2 ; chlorin, chlorophylls a, b, d, f and bacteriochlorophylls c, d, e, f; and bacteriochlorin, bacteriochlorophylls a, b, g. The spectra are compared with those of the same pigment in liquid solution (predominantly 2-methyltetrahydrofuran) at room temperature (293 K). The measured Stokes shifts at 77 K across the 12 macrocycles range from ~30 to 300 cm −1 . The spectral data in digital form are made available as part of the PhotochemCAD databases. Literature searches have revealed extensive published data for Chl a (often in biological matrices) but at best rather limited data for less common macrocycles. The availability of a systematic collection of curated spectral data collected at low temperature should be useful for a variety of assessments, including reconstruction of absorption spectra of (bacterio)chlorophyll-containing protein complexes, vibrational analysis of absorption and fluorescence spectra, and calculations where knowledge of energy levels is important.

Niedzwiedzki, Dariusz M. [Washington University in↗

An automated integrated web-based smart tool for open stope design

The Stability Graph is a widely used tool for the design of open stopes in underground mining. Many users of the Stability Graph still apply this design method manually. Although the manual approach has benefits, using multiple graphs and stability number computation charts for each stope surface is time-consuming, even for the experienced mining engineer. Current practice in the use of the method also limits data sharing. This paper presents a StopeSoft web-based tool for open stope stability prediction that is developed on the basis of the Stability Graph method and is available at openstope.com. StopeSoft incorporates flexibility in terms of Stability Graph options and incorporates additional critical factors often overlooked. As a web-based tool, StopeSoft encourages and makes data sharing possible globally, focused on expanding the database and improving the current limitations of the Stability Graph to provide practical, reliable solutions for mining engineers, consultants, and academics. The StopeSoft automated process facilitates the process of open stope stability prediction, saving time and minimizing potential human errors. Statistical treatment of the data accounts for the variability of input parameters to emphasize the probabilistic nature of the Stability Graph method. The probabilistic interpretation of the stability states of stope surfaces eliminates the false feeling of absolute stope performance based on its location on the Stability Graph , as implied by the deterministic approach.

58 GEOSCIENCES↗

Sharing is caring: An extensive analysis of parameter-based transfer learning for the prediction of building thermal dynamics

In recent years deep neural networks have been proposed as a lightweight data-driven model to capture high-dimensional, nonlinear physical processes to predict building thermal responses. However, the need of a large amount of data for the training process of deep neural networks clashes with the potential limited data availability in most existing or new buildings. Transfer learning aims to enhance the performance of a target learner exploiting knowledge from related and similar environments. This study conducted a suite of experiments that leveraged 250 data-driven models based on a synthetic dataset of a building archetype to study the influence of data availability, energy efficiency level, occupancy and climate for the transfer process of thermal dynamics. The performance of the transfer learning process was compared against a classical machine learning approach. Here, the results suggest that building thermal dynamics can be effectively transferred under the same climatic conditions, increasing performance when dealing with different occupancy schedules, efficiency levels and low data availability. Furthermore, the paper compares the performance of both transfer learning and machine learning approaches in an online fashion, to support the implementation in real-world deployment.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Identifying Adversarial Cyber-Activity in Operational Technology Environments Using Bayesian Networks

Critical infrastructure and other operational technology (OT) environments face increasing cybersecurity risks from adversarial behavior. This paper describes the development of a risk model using a Bayesian network to enhance the comprehension of observable cyber events caused by malicious activity in OT environments. The core of the Bayesian network is a process model that describes the stages of adversary behavior. The remainder of the model is based on the MITRE ATT&CK® for Industrial Control Systems (ICS) taxonomy, which includes tactics and techniques that may be used by the adversary. The observables provide evidence for adversary behavior through the intermediary technique and tactic nodes. One challenge in constructing this model is a lack of open-source data from cyber-attacks on OT systems. This paper discusses learning from limited data, the elicitation of expert opinion to construct the conditional probability tables when data is scarce, and the refinement of the most difficult conditional probabilities tables using several forms of sensitivity analyses. Finally, the Bayesian network is demonstrated using two historical case studies: the DarkSide ransomware attack on the Colonial Pipeline and the destructive cyberattack targeting the ThyssenKrupp blast furnace. Index Terms—Cybersecurity, industrial control systems, operational technology

97 - MATHEMATICS AND COMPUTING↗

Residuals-based distributionally robust optimization with covariate information

We consider data-driven approaches that integrate a machine learning prediction model within distributionally robust optimization (DRO) given limited joint observations of uncertain parameters and covariates. Our framework is flexible in the sense that it can accommodate a variety of regression setups and DRO ambiguity sets. We investigate asymptotic and finite sample properties of solutions obtained using Wasserstein, sample robust optimization, and phi-divergence-based ambiguity sets within our DRO formulations, and explore cross-validation approaches for sizing these ambiguity sets. Through numerical experiments, we validate our theoretical results, study the effectiveness of our approaches for sizing ambiguity sets, and illustrate the benefits of our DRO formulations in the limited data regime even when the prediction model is misspecified.

97 MATHEMATICS AND COMPUTING↗

Residuals-based distributionally robust optimization with covariate information

We consider data-driven approaches that integrate a machine learning prediction model within distributionally robust optimization (DRO) given limited joint observations of uncertain parameters and covariates. Our framework is flexible in the sense that it can accommodate a variety of regression setups and DRO ambiguity sets. We investigate asymptotic and finite sample properties of solutions obtained using Wasserstein, sample robust optimization, and phi-divergence-based ambiguity sets within our DRO formulations, and explore cross-validation approaches for sizing these ambiguity sets. Through numerical experiments, we validate our theoretical results, study the effectiveness of our approaches for sizing ambiguity sets, and illustrate the benefits of our DRO formulations in the limited data regime even when the prediction model is misspecified.

97 MATHEMATICS AND COMPUTING↗

47 Tuc in Rubin Data Preview 1. Exploring Early LSST Data and Science Potential

We present analyses of the early data from Rubin Observatory’s Data Preview 1 (DP1) for the field of the globular cluster 47 Tuc. The DP1 data set for 47 Tuc includes four nights of observations from the Rubin Commissioning Camera (LSSTComCam), covering multiple bands (ugriy). We address challenges of crowding in the inner region of the cluster and toward the SMC in DP1, and demonstrate improved star–galaxy separation by fitting fifth-degree polynomials to the stellar loci in color–color diagrams and applying multidimensional sigma clipping. We compile a catalog of 3576 probable 47 Tuc member stars selected via a combination of isochrone, Gaia proper-motion, and color–color space matched filtering. We explore the sources of photometric scatter in the 47 Tuc color–color sequence, evaluating contributions from various potential sources, including differential extinction within the cluster. Finally, of the 72 well-characterized variables in the field, we recover three known variable stars, including two RR Lyrae and one eclipsing binary, in the coadd-based object catalog, and identify 62 in the difference image-based object catalog. Although the DP1 lightcurves have sparse temporal sampling, they appear to follow the patterns of densely sampled literature lightcurves well. Despite some data limitations for crowded-field stellar analysis, DP1 demonstrates the promising scientific potential for future LSST data releases.

Choi, Yumi [NSF National Optical-Infrared Astronom↗

Extrapolation of the Rainflow-Counted Load Ranges for Fatigue Assessment of the Wind Turbine's Blades

Wind turbine design standards recommend the use of statistical modeling coupled with extrapolation of the short-term load data to long-term periods for fatigue reliability assessment. However, statistical error and computational expense can limit the accuracy of such approaches. In the case of wind turbine blades, the errors are more significant because of the high material fatigue exponent that makes the damage estimations more sensitive to variations. In addition, due to different excitation sources, the flapwise load range histogram is not unimodal, and thus its statistical modeling is complex. In the present work, we provide three methods for statistical modeling of the flapwise bending moment ranges including a novel approach based on frequency-based separation of the modes. The first two methods are simplified approaches for modeling the most crucial load ranges using unimodal distributions and the third method involves multimodal distribution fitting. The research is based on 3600 10-minute aeroelastic simulations of DTU 10MW case study wind turbine from which a benchmark damage equivalent load (DEL) is calculated. The DEL calculated by each of the three proposed methods is compared to this reference. The results show that the conventional approach based on using 6 seeds as well as using mixture models fitted on the limited data lead to under-conservative results with errors up to 23%. On the other hand, the simplified unimodal approaches provided in this work can provide conservative estimations of the fatigue damage with mean values 5% and 12% higher than the benchmark. However, the variability of the DEL estimates is higher when using unimodal extrapolation of the load ranges, and the data can be conservative by 17.5%. The proposed unimodal fits suggested for modeling and extrapolation of the blade's load ranges provide less errors relatively and most importantly conservative DEL estimations while maintaining computational efficiency.

blade fatigue↗

Direct Feed High-Level Waste APPS Model Glass Testing (DFHLW APPS) Matrix, Phase 2

This report summarizes the data collected during the batching and melting of a second matrix of Direct Feed High-Level Waste (DFHLW) glasses generated using the preliminary enhanced waste glass models (EWG2.5) and the Britton and Anderson (2024) preliminary DFHLW feed vector. The purpose of these glasses is two-fold: 1. Validate EWG2.5 glass calculations being used in the Aspen Process Performance Simulation (APPS) model. 2. Evaluate and ultimately improve the glass property models and formulation methods used for design of DFHLW glasses as part of an iterative process of data collection and model refinement. Some of the 16 APPS2 glasses tested did not satisfy all target property constraints due to the limited data on DFHLW glass supporting the EWG2.5 models. • One glass, APPS2-10, formed nepheline on canister centerline cooling (CCC) heat-treatment and failed the product consistency test (PCT) response limits. This glass also had high B and Cr release rates for the toxicity characteristic leaching procedure (TCLP). All other glasses were found to satisfy the PCT and TCLP constraints for both quenched and CCC samples. • One glass, APPS2-08, had higher than acceptable viscosity due to magnetite crystallization. • One glass, APPS2-09, formed greater than 2 vol% crystals at 950 °C. As the glass design criterion was that the temperature at 2 vol% crystal (T 2% ) be less than 950 °C, only one glass failed the criteria. However, this criterion is being reevaluated. Four additional glasses formed crystal fractions between 1 and 2 vol% at 950 °C (APPS2-03, -08, -12, and -14). • Four glasses – APPS2-01, -02, -04, and -16 – failed the Monofrax K-3 refractory neck corrosion (k neck ) design limit of 0.04 in. at 1208 °C for 6 d. This is another criterion being reevaluated. Four additional glasses (APPS2-05, -06, -11, and -13) exhibited 0.025 = k neck = 0.04 in. • All 16 glasses passed the sulfur solubility and TCLP constraints. The measured property values were compared to predicted values using EWG2.5 and a selection of other existing models. A few models (e.g., electrical conductivity, TCLP) were found to be adequate for designing DFHLW glasses in the near future, while others require refits or offsets. It is recommended that new property models be developed for EWG3.0, as a large amount of DFHLW glass property data (> 14 × existing data) is expected to be collected in the compositional spaces where no data was previously available. To enable near-term calculations and formulations for designing DFHLW glasses and processing rate estimations, a formulation algorithm with minor modifications will be developed, EWG2.6.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

A versatile high-speed x-ray microscope for sub-10 nm imaging

We have developed a next-generation scanning x-ray microscope RASMI (RApid Scanning Microscopy Instrument) for high-throughput tomographic imaging. RASMI is installed at the hard x-ray nanoprobe beamline at NSLS-II and is capable of manipulating 1D multilayer Laue lenses (MLLs) and 2D optics (both zone plates and monolithically assembled 2D MLLs). The sample scanning stage utilizes line-focusing interferometry as an encoder while performing fly-scanning data acquisition. The system can be configured for both position- and time-triggering modes during fly-scanning. The microscope demonstrated a detector-limited data acquisition rate of 1.25 kHz during ptychography measurements. The initial x-ray results yielded a sample-limited resolution of ∼6 nm in 2D. RASMI can be adopted for in-vacuum applications and is a foundation for the next-generation scanning microscopy systems to be developed and commissioned at NSLS-II.

36 MATERIALS SCIENCE↗

Conceptual Designs for Irradiation Creep Testing of SiC in HFIR

Understanding irradiation creep of nuclear fuel cladding is important to properly size the initial fuel-cladding gap and understand when pellet-cladding contact is expected to occur due to a combination of fuel swelling and cladding creep-down. Irradiation creep also plays a role in relaxing stresses that develop in-pile. Silicon carbide fiber–reinforced silicon carbide matrix (SiC/SiC) composites are the leading long-term accident-tolerant fuel cladding concept for light-water reactors (LWRs). Although some limited data are available regarding irradiation creep of the individual constituents (fibers, matrix), data regarding irradiation creep of SiC/SiC composites are currently insufficient. Additional data regarding irradiation creep compliance and the rupture lifetime (combination of creep and slow crack growth) are needed to understand material limitations. This work describes the design and development of two irradiation vehicles that are being pursued for testing SiC/SiC concepts in the High Flux Isotope Reactor (HFIR). The first is a passive experiment, referred to as the PRECISE experiment, that leverages the constant coolant pressure of HFIR to compress a metallic bellows and provide a well-characterized load to drive creep in a SiC/SiC dog bone specimen. The total creep strain would be quantified post-irradiation by measuring dimensional changes of the specimen length as well as local dimensional changes within the gauge region. Non-stressed specimens would also be irradiated under the same conditions to provide an indication of dimensional changes due to radiation-induced swelling in the absence of creep. A second, more complex experiment, referred to as the INSITE experiment, is being designed in parallel that would use pneumatics to pressurize a metal bellows and linear variable differential transformers (LVDTs) to measure the specimen displacement in situ during irradiation. Such an experiment would provide significantly more data regarding the evolution of the creep compliance as a function of dose and applied stress within a single experiment but would require significantly more development time and cost to execute. The primary concern with the INSITE experiment is the accuracy, reliability, and expected lifetime of the LVDTs during irradiation at elevated temperatures. Efforts are being made to adjust the experiment design and operating procedure to limit LVDT temperatures and mitigate or otherwise compensate for uncertainties due to factors such as temperature fluctuations, creep in the surrounding structural materials, and drift of the LVDTs. This work describes the experiment designs, thermal and structural analysis that were performed to ensure that the desired temperature and stress conditions can be achieved, some initial sensitivity analyses to predict the evolution of the radiation-induced specimen displacements, and potential sources of uncertainty in the measurements. Out-of-pile testing is being performed in parallel to confirm that the test trains achieve the expected stress states in the specimens and do not result in prohibitive stress concentrators (e.g., in the grip regions) that might risk pre-mature failure. The PRECISE experiments are proceeding toward fabrication and assembly with HFIR insertion planned during fiscal year 2026. The INSITE experiment is progressing toward out-of-pile demonstrations, which will provide more conclusive evidence regarding the feasibility of executing these tests in HFIR or whether alternative displacement monitoring techniques may need to be considered.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Dynamic Boundary Microgrids Under Privatization Considerations

Microgrids have physical, electrical, and logical (data, network, and ownership) boundaries. To power unserved customer loads during an outage, microgrids can extend the traditional operational boundaries. This can become complex when considering microgrid-to-microgrid (M2M) interactions where sensitive information such as competitive microgrid operational data is not shared. This work proposes an optimization method coordinated between microgrid controllers and distribution management systems that limits data sharing. The method involves a competitive bidding strategy that maximizes unserved load coverage while minimizing resource utilization and sensitive operational data sharing among entities. The work is validated on a two-microgrid system with photovoltaic and energy storage systems and curves of load derived from real world residential buildings datasets. Results show that the proposed method, when applied for three distinct use cases of energy storage sufficiency to cover the predefined boundary and/or the expanded boundary, can successfully select and bid the available load coverage.

Starke, Michael [ORNL] (ORCID:0000000221211195)↗

Baseline Cost Model for Hydropower: Documentation (2025)

Hydropower currently contributes about 80 GW of conventional and 23 GW of pumped storage capacity to the United States (US) power grid. Previous studies have estimated a considerable amount of remaining US hydropower resources, including non-powered dams (NPD) (Hadjerioua et al., 2012), new stream-reach developments (NSD) (Kao et al., 2014), pumped storage hydropower (PSH) and canal/conduit (Kao et al., 2022). The combined theoretical capacity potential of these various hydropower resources is comparable to the existing US hydropower capacity. There is a continuing interest in developing this hydropower potential, particularly to help meet the increasing demand for electricity. However, available data (Sasthav and Oladosu, 2022) show that the rate of new hydropower development has slowed considerably over time despite the interest of industry stakeholders. This is partly due to the competition from other energy resources and from the highly dispersed nature of remaining hydropower resources, which lead to high information requirements for evaluating the feasibility of potential projects. Cost information provides the most succinct summary of the feasibility of a potential hydropower project required by stakeholders, including developers, investors, policymakers, consumer groups, etc., considering investment options. The best estimates of hydropower costs can be obtained through detailed engineering design and cost assessments of individual projects. However, this approach has high data and resource (time, funds, cross-disciplinary expertise) requirements that render it inapplicable for rapid cost estimation with limited data. Although innovative approaches can overcome some of these impediments (see Oladosu and Ma, 2024 for such an application to potential NPD projects), the development of such approaches still requires significant amounts of resources and are not generally applicable to all hydropower project types. Therefore, statistical and parametric methods using simpler cost specifications remain of significant utility to hydropower stakeholders and are, at the least, complementary to more detailed approaches, particularly when evaluating many potential projects.

13 HYDRO ENERGY↗

Integrated Framework of Multisource Data Fusion for Outage Location in Looped Distribution Systems

Accurate outage location is essential for expediting post-outage power restoration, minimizing outage duration, and enhancing the resilience of distribution networks. With the advent of advanced metering infrastructure, data-driven outage location methods have significantly advanced beyond traditional approaches that rely on manual inspections. However, existing methods still face critical challenges, like reliance on single-source data, limited ability to handle partially observable systems or difficulties with loop networks. To the best of our knowledge, no single approach has comprehensively addressed all of these challenges at once. To this end, this paper proposes a comprehensive multisource data fusion framework for outage locations via probabilistic graph networks. The framework consists of three key phases. First, a novel method for reconstituting distribution networks with loops is developed, transforming looped networks into multiple radial subnetworks that retain all outage causalities of the original network. Second, Bayesian network (BN) models are established for each subnetwork, integrating multiple data sources and network structures. Finally, a joint Gibbs sampling mechanism, featuring forward and backward information flow, is designed to merge data from separate BN models and maximize the utilization of limited evidence, ensuring accurate outage location identification. In conclusion, the framework was validated on two modified public test systems, and comparative studies confirmed its effectiveness.

24 POWER TRANSMISSION AND DISTRIBUTION↗

The Design and Implementation of a Secure Datastore Based on Ethereum Smart Contract

In this paper, we present a secure datastore based on an Ethereum smart contract. Our research is guided by three research questions. First, we will explore to what extend a smart-contract-based datastore should resemble a traditional database system. Second, we will investigate how to store the data in a smart-contract-based datastore for maximum flexibility while minimizing the gas consumption. Third, we seek answers regarding whether or not a smart-contract-based datastore should incorporate complex processing such as data encryption and data analytic algorithms. The proposed smart-contract-based datastore aims to strike a good balance between several constraints: (1) smart contracts are publicly visible, which may create a confidentiality concern for the data stored in the datastore; (2) unlike traditional database systems, the Ethereum smart contract programming language (i.e., Solidity) offers very limited data structures for data management; (3) all operations that mutate the blockchain state would incur financial costs and the developers for smart contracts must make sure sufficient gas is provisioned for every smart contract call, and ideally, the gas consumption should be minimized. Our investigation shows that although it is essential for a smart-contract-based datastore to offer some basic data query functionality, it is impractical to offer query flexibility that resembles that of a traditional database system. Furthermore, we propose that data should be structured as tag-value pairs, where the tag serves as a non-unique key that describes the nature of the value. We also conclude that complex processing should not be allowed in the smart contract due to the financial burden and security concerns. The tag-based secure datastore designed this way also defines its applicative perimeter, i.e., only applications that align with our strategy would find the proposed datastore a good fit. Those that would rather incur higher financial cost for more data query flexibility and/or less user burden on data pre- and post-processing would find the proposed database too restrictive.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗