Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Data Sources”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 127 records · Page 7

Electric Vehicle Charging Analytics and Reporting Tool (EV-ChART): Data Format and Preparation Guidance, Version 2.0

The Joint Office of Energy and Transportation maintains the Electric Vehicle Charging Analytics and Reporting Tool (EV-ChART), which provides a centralized hub for submitting electric vehicle (EV) charging infrastructure data directed by the Federal Highway Administration (23 CFR 680.112) EV-ChART will provide a streamlined data submission process and an integrated set of analytic tools, connect to other data sources, and empower data sharing and access across stakeholders, including the public. Any data shared publicly will be aggregated and anonymized to stay in accordance with 23 CFR 680. This EV-ChART Data Format and Preparation Guidance provides a comprehensive overview of the data reporting requirements as authorized under 23 CFR 680.112. The guidance is intended to be used alongside the EV-ChART Data Input Template, which defines the tabular data structure that these data submissions must follow.

ADVANCED PROPULSION SYSTEMS,MATHEMATICS AND COMPUT↗

VA EDH Data Curation Documentation FY25-Q1

This data source documentation report provides researchers with valuable insights into the structure, contents, and data sources used to compile the datasets. It specifically covers the Fiscal Year 2024, Fourth Quarter (FY25-Q1) dataset curation documentation for the Environmental Determinants of Health (EDH) project.

97 MATHEMATICS AND COMPUTING↗

Deep Learning for Full Waveform Inversion of Elastic Active-Source Seismic Data to Estimate P-Wave Velocity Models

Seismic imaging methods are critical for Global Security and Energy & Homeland Security missions and activities that rely on subsurface characterization, but traditional methods remain computationally expensive and require significant labor hours and expertise to execute. Within the past few years, machine learning (ML), namely deep learning (DL), has been used to develop data-driven end-to-end full waveform inversion (FWI) methods to estimate 2D P-wave velocity (Vp) models in a fraction of the time as conventional FWI. These methods, however, are trained on simplistic acoustic wave seismic data and Vp models that are not realistic nor representative of real-world observations, leaving a large gap between the state-of-the-art and deployable, feasible, and practical DL FWI methods. Here, we generate a synthetic active-source, 3D, elastic wave seismic data set and a variety of Vp models with realistic geologic structure for training DL FWI methods. We evaluate six different methods that have performed well for acoustic DL FWI or medical imaging tasks using our more realistic dataset. We find that these six trained models do not match the performance of published acoustic end-to-end DL FWI methods, indicating more training data may be needed, physics may need to be incorporated to achieve good accuracy at the sacrifice of the end-to-end advantage, and/or novel methods need to be developed to enable end-to-end DL FWI methods to perform well for real-world seismic data.

45 MILITARY TECHNOLOGY, WEAPONRY, AND NATIONAL DEF↗

The Coastal Carbon Library and Atlas: Open source soil data and tools supporting blue carbon research and policy

Abstract Quantifying carbon fluxes into and out of coastal soils is critical to meeting greenhouse gas reduction and coastal resiliency goals. Numerous ‘blue carbon’ studies have generated, or benefitted from, synthetic datasets. However, the community those efforts inspired does not have a centralized, standardized database of disaggregated data used to estimate carbon stocks and fluxes. In this paper, we describe a data structure designed to standardize data reporting, maximize reuse, and maintain a chain of credit from synthesis to original source. We introduce version 1.0.0. of the Coastal Carbon Library, a global database of 6723 soil profiles representing blue carbon‐storing systems including marshes, mangroves, tidal freshwater forests, and seagrasses. We also present the Coastal Carbon Atlas, an R‐shiny application that can be used to visualize, query, and download portions of the Coastal Carbon Library. The majority (4815) of entries in the database can be used for carbon stock assessments without the need for interpolating missing soil variables, 533 are available for estimating carbon burial rate, and 326 are useful for fitting dynamic soil formation models. Organic matter density significantly varied by habitat with tidal freshwater forests having the highest density, and seagrasses having the lowest. Future work could involve expansion of the synthesis to include more deep stock assessments, increasing the representation of data outside of the U.S., and increasing the amount of data available for mangroves and seagrasses, especially carbon burial rate data. We present proposed best practices for blue carbon data including an emphasis on disaggregation, data publication, dataset documentation, and use of standardized vocabulary and templates whenever appropriate. To conclude, the Coastal Carbon Library and Atlas serve as a general example of a grassroots F.A.I.R. (Findable, Accessible, Interoperable, and Reusable) data effort demonstrating how data producers can coordinate to develop tools relevant to policy and decision‐making.

Holmquist, James R.↗

Global nitrogen deposition inputs to cropland at national scale from 1961 to 2020

Nitrogen (N) deposition is a significant nutrient input to cropland and consequently important for the evaluation of N budgets and N use efficiency (NUE) at different scales and over time. However, the spatiotemporal coverage of N deposition measurements is limited globally, whereas modeled N deposition values carry uncertainties. Here, we reviewed existing methods and related data sources for quantifying N deposition inputs to crop production on a national scale. We utilized different data sources to estimate N deposition input to crop production at national scale and compared our estimates with 14 N budget datasets, as well as measured N deposition data from observation networks in 9 countries. We created four datasets of N deposition inputs on cropland during 1961–2020 for 236 countries. These products showed good agreement for the majority of countries and can be used in the modeling and assessment of NUE at national and global scales. One of the datasets is recommended for general use in regional to global N budget and NUE estimates.

54 ENVIRONMENTAL SCIENCES↗

SEEDling [SWR-20-99]

SEEDling is intended to simplify the merging of multiple data sources prior to importing into the SEED platform. Geographic information about buildings and taxlots is often obtained from various data sources, and the process of synthesizing this information for many buildings across a city can be burdensome for city data managers. SEEDling uses GIS tools to automate and simplify this process, as well as apply criteria for SEED inclusion (ex: sqft > 50k) and assign UBIDs, to create records that can be easily imported into the SEED platform.

Long, Nicholas↗

Development of a Geothermal Module in reV: Quantifying the Geothermal Potential while Accounting for the Geospatial Intersection of the Grid Infrastructure and Land Use Characteristics

The Renewable Energy Potential (reV) model is a geospatial platform for estimating technical potential and developing renewable energy supply curves, initially developed for wind and solar technologies. The model evaluates deployment constraints, considering land use, environmental, and cultural factors, and estimates the distance to existing grid features to connect future plants (Maclaurin et al., 2021). A pressing deficiency in the reV model, however, is representation of geothermal electricity generation technologies. To address this gap, we developed a novel geothermal generation module for reV that allows for representation and analysis at the same level of detail as other renewable technologies. This paper describes our process for evaluating data sources for the modeling, and presents five preliminary reV geothermal results. More specifically, we present two sets of resource data that represent upper and lower bounds for geothermal potential. We then present several sensitivity runs using the upper bound resource data; the results are encouraging that levelized cost of electricity (LCOE) can be reduced by optimizing the location and estimated capacity of the spatially diverse geothermal resource while considering the distance to existing grid infrastructure. Our preliminary supply curves and levelized cost of electricity (LCOE) results should be considered with care due to the highly uncertainty in geothermal resource potential data. We present median LCOE values for the conterminous U.S. for five scenarios: four hydrothermal (3.5km depth) and one EGS (4.5km depth). The capital and operating costs for each respective technology are modeled. We also compare results using two different resource data sources.

exclusions↗

Model form and sensitivity analysis of CALPHAD-based nucleation models in b-stabilized Ti alloys

Accurate prediction of α-phase nucleation and growth in β-stabilized titanium alloys is crucial for designing heat treatments to optimize mechanical properties in additively manufactured lightweight components. Ideally, predictions of nucleation and growth would incorporate both top-down observations of past experimental heat treatments and bottom-up modeling of phase transformations; however, the appropriate method of combining these information sources is not self-evident. Combining top-down and bottom-up information requires a unified form of model that can connect between spatiotemporal scales, as well as sets of fitting parameters that can be identified by each data source. The selection of which parameters to fit to which data source can be made based on expert opinion, or by performing a sensitivity analysis. In solid-solid nucleation, direct observation of the nucleation and growth process is challenging. Most data on the heat treatment-controlled phase transformations are not in-situ. To predict the process and outcome of the nucleation, growth and coarsening of precipitates, theoretical models of the nucleation pathway are used to bridge the gap. Many sources of uncertainty affect the modeling of this nucleation process. It can be influenced by small variations in the thermomechanical processing history, chemical composition, and initial microstructure. If molecular dynamics (MD) simulations are used to determine thermodynamic quantities and inform CALPHAD modeling, additional uncertainty can be introduced and accounted for using Bayesian methods. Top-down uncertainties require additional steps to quantify. The influence of nucleation model form on the sensitivity of predictions to input parameters and physical conditions is the focus of this study. Classical nucleation theory (CNT) allows modeling to formulate the nucleation as homogeneous or, more commonly, heterogeneous. Non-classical nucleation models are also increasingly explored as a means of reconciling top-down and bottom-up data. In this study, the sensitivity of the intragranular nucleation of α in a β-annealed, slow-cooled aging (BASCA) heat treatment of β-stabilized Ti5553 alloy is explored using CNT and both heterogeneous and homogeneous assumptions. The Kampmann-Wagner Numerical model of precipitate nucleation and growth is employed. Using open-source tools (pyCalphad and thermodynamic modeling of TiMo as a surrogate system, a sensitivity analysis is performed to measure variations in key parameters, including chemical driving force, interfacial energy, and diffusivity, as they relate to predictions of precipitate number density. The inclusion of top-down and bottom-up data in selection of nucleation model form is discussed.

Rodriguez Negron, A. M.↗

A causal data fusion method for the general exposure and outcome

Abstract With the advent of the big data era, the need to combine multiple individual data sets to draw causal effects arises naturally in many medical and biological applications. Especially each data set cannot measure enough confounders to infer the causal effect of an exposure on an outcome. In this article, we extend the method proposed by a previous study to causal data fusion of more than two data sets without external validation and to a more general (continuous or discrete) exposure and outcome. Theoretically, we obtain the condition for identifiability of exposure effects using multiple individual data sources for the continuous or discrete exposure and outcome. The simulation results show that our proposed causal data fusion method has unbiased causal effect estimate and higher precision than traditional regression, meta‐analysis and statistical matching methods. We further apply our method to study the causal effect of BMI on glucose level in individuals with diabetes by combining two data sets. Our method is essential for causal data fusion and provides important insights into the ongoing discourse on the empirical analysis of merging multiple individual data sources.

Li, Hongkai↗

Development of a Geothermal Module in reV: Quantifying the Geothermal Potential While Accounting for the Geospatial Intersection of the Grid Infrastructure and Land Use Characteristics: Preprint

The Renewable Energy Potential (reV) model is a geospatial platform for estimating technical potential and developing renewable energy supply curves, initially developed for wind and solar technologies. The model evaluates deployment constraints, considering land use, environmental, and cultural factors, and estimates the distance to existing grid features to connect future plants (Maclaurin et al., 2021). A pressing deficiency in the reV model, however, is representation of geothermal electricity generation technologies. To address this gap, we developed a novel geothermal generation module for reV that allows for representation and analysis at the same level of detail as other renewable technologies. This paper describes our process for evaluating data sources for the modeling, and presents five preliminary reV geothermal results. More specifically, we present two sets of resource data that represent upper and lower bounds for geothermal potential. We then present several sensitivity runs using the upper bound resource data; the results are encouraging that levelized cost of electricity (LCOE) can be reduced by optimizing the location and estimated capacity of the spatially diverse geothermal resource while considering the distance to existing grid infrastructure. Our preliminary supply curves and levelized cost of electricity (LCOE) results should be considered with care due to the highly uncertainty in geothermal resource potential data. We present median LCOE values for the conterminous U.S. for five scenarios: four hydrothermal (3.5km depth) and one EGS (4.5km depth). The capital and operating costs for each respective technology are modeled. We also compare results using two different resource data sources.

exclusions↗

Cloud-based Testbed for Adaptive Under-Frequency Load Shedding with High DER Penetration

Increasing penetration of distributed energy resources and behind-the-meter renewables may soon disrupt the efficacy of critical protection schemes, such as under-frequency load shedding (UFLS). Improved data exchange and coordination across the transmission-distribution boundary will be required to maintain reliability of bulk electric system. Standards-based data integration platforms using agreed-upon semantic vocabularies, such as the Common Information Model, will be key to enabling adaptive protection schemes requiring synthesized data from both the bulk power system and behind-the-meter resources. This paper introduces a cloud-based open-source data integration environment and UFLS clustering algorithm being developed to enable adaptive relay coordination between transmission and distribution utilities in the state of Vermont.

Anderson, Alexander A.↗

Multitask methods for predicting molecular properties from heterogeneous data

Data generation remains a bottleneck in training surrogate models to predict molecular properties. We demonstrate that multitask Gaussian process regression overcomes this limitation by leveraging both expensive and cheap data sources. In particular, we consider training sets constructed from coupled-cluster (CC) and density functional theory (DFT) data. We report that multitask surrogates can predict at CC-level accuracy with a reduction in data generation cost by over an order of magnitude. Of note, our approach allows the training set to include DFT data generated by a heterogeneous mix of exchange–correlation functionals without imposing any artificial hierarchy on functional accuracy. More generally, the multitask framework can accommodate a wider range of training set structures—including the full disparity between the different levels of fidelity—than existing kernel approaches based on Δ-learning although we show that the accuracy of the two approaches can be similar. Consequently, multitask regression can be a tool for reducing data generation costs even further by opportunistically exploiting existing data sources.

Chemistry↗

Renewable Energy Potential Model: Geothermal Supply Curves

The Renewable Energy Potential (reV) model is a geospatial platform for estimating technical potential and developing renewable energy supply curves, initially developed for wind and solar technologies. The model evaluates deployment constraints, considering land use, environmental, and cultural factors, and estimates the distance to existing grid features to connect future plants (Maclaurin et al., 2021). A pressing deficiency in the reV model, however, is representation of geothermal electricity generation technologies. To address this gap, we developed a novel geothermal generation module for reV that allows for representation and analysis at the same level of detail as other renewable technologies. The included paper describes our process for evaluating data sources for the modeling, and presents five preliminary reV geothermal results. More specifically, we present two sets of resource data that represent upper and lower bounds for geothermal potential. We then present several sensitivity runs using the upper bound resource data; the results are encouraging that levelized cost of electricity (LCOE) can be reduced by optimizing the location and estimated capacity of the spatially diverse geothermal resource while considering the distance to existing grid infrastructure. Our preliminary supply curves and levelized cost of electricity (LCOE) results provided here should be considered with care due to the high uncertainty in geothermal resource potential data. We present median LCOE values for the conterminous U.S. for three scenarios: two hydrothermal (3.5km depth, USGS heat flow & SMU temperatures respectively) and one EGS (4.5km depth, SMU temperatures). The capital and operating costs for each respective technology are modeled. We also compare results using two different resource data sources.

15 GEOTHERMAL ENERGY↗

Integrated Framework of Multisource Data Fusion for Outage Location in Looped Distribution Systems

Accurate outage location is essential for expediting post-outage power restoration, minimizing outage duration, and enhancing the resilience of distribution networks. With the advent of advanced metering infrastructure, data-driven outage location methods have significantly advanced beyond traditional approaches that rely on manual inspections. However, existing methods still face critical challenges, like reliance on single-source data, limited ability to handle partially observable systems or difficulties with loop networks. To the best of our knowledge, no single approach has comprehensively addressed all of these challenges at once. To this end, this paper proposes a comprehensive multisource data fusion framework for outage locations via probabilistic graph networks. The framework consists of three key phases. First, a novel method for reconstituting distribution networks with loops is developed, transforming looped networks into multiple radial subnetworks that retain all outage causalities of the original network. Second, Bayesian network (BN) models are established for each subnetwork, integrating multiple data sources and network structures. Finally, a joint Gibbs sampling mechanism, featuring forward and backward information flow, is designed to merge data from separate BN models and maximize the utilization of limited evidence, ensuring accurate outage location identification. In conclusion, the framework was validated on two modified public test systems, and comparative studies confirmed its effectiveness.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Chimera D-Series Gravitational Wave Emission Sourced from Matter

Gravitational wave data sourced from the the time-dependent fluid quadrupole motion, in the Chimera D-Series three-dimensional core collapse supernova simulations. Data from three models initiated from three different progenitors are presented: D9.6-3D, D15-3D, and D25-3D. Please see the README for more information about the data structure and progenitors.

79 ASTRONOMY AND ASTROPHYSICS↗

The United States COVID-19 Forecast Hub dataset

Academic researchers, government agencies, industry groups, and individuals have produced forecasts at an unprecedented scale during the COVID-19 pandemic. To leverage these forecasts, the United States Centers for Disease Control and Prevention (CDC) partnered with an academic research lab at the University of Massachusetts Amherst to create the US COVID-19 Forecast Hub. Launched in April 2020, the Forecast Hub is a dataset with point and probabilistic forecasts of incident cases, incident hospitalizations, incident deaths, and cumulative deaths due to COVID-19 at county, state, and national, levels in the United States. Included forecasts represent a variety of modeling approaches, data sources, and assumptions regarding the spread of COVID-19. The goal of this dataset is to establish a standardized and comparable set of short-term forecasts from modeling teams. These data can be used to develop ensemble models, communicate forecasts to the public, create visualizations, compare models, and inform policies regarding COVID-19 mitigation. These open-source data are available via download from GitHub, through an online API, and through R packages.

60 APPLIED LIFE SCIENCES↗

HarDWR - Harmonized Water Rights Records

A dataset within the Harmonized Database of Western U.S. Water Rights (HarDWR). For a detailed description of the database, please see the meta-record v2.0. Changelog v2.0 - Recalculated based on data sourced from WestDAAT - Changed using a Site ID column to identify unique records to using aa combination of Site ID and Allocation ID - Removed the Water Management Area (WMA) column from the harmonized records. The replacement is a separate file which stores the relationship between allocations and WMAs. This allows for allocations to contribute to water right amounts to multiple WMAs during the subsequent cumulative process. - Added a column describing a water rights legal status - Added "Unspecified" was a water source category - Added an acre-foot (AF) column - Added a column for the classification of the right's owner v1.02 - Added a .RData file to the dataset as a convenience for anyone exploring our code. This is an internal file, and the one referenced in analysis scripts as the data objects are already in R data objects. v1.01 - Updated the names of each file with an ID number less than 3 digits to include leading 0s v1.0 - Initial public release Description Here we present an updated database of Western U.S. water right records. This database provides consistent unique identifiers for each water right record, and a consistent categorization scheme that puts each water right record into one of seven broad use categories. These data were instrumental in conducting a study of the multi-sector dynamics of inter-sectoral water allocation changes though water markets (Grogan et al., *in review*). Specifically, the data were formatted for use as input to a process-based hydrologic model, Water Balance Model (WBM), with a water rights module (Grogan et al., *in review*). While this specific study motivated the development of the database presented here, water management in the U.S. West is a rich area of study (e.g., Anderson and Woosly, 2005; Tidwell, 2014; Null and Prudencio, 2016; Carney et al., 2021) so releasing this database publicly with documentation and usage notes will enable other researchers to do further work on water management in the U.S. West. We produced the water rights database presented here in four main steps: (1) data collection, (2) data quality control, (3) data harmonization, and (4) generation of cumulative water rights curves. Each of steps (1)-(3) had to be completed in order to produce (4), the final product that was used in the modeling exercise in Grogan et al. (*in review*). All data in each step is associated with a spatial unit called a Water Management Area (WMA), which is the unit of water right administration utilized by the state in which the right came from. Steps (2) and (3) required use to make assumptions and interpretation, and to remove records from the raw data collection. We describe each of these assumptions and interpretations below so that other researchers can choose to implement alternative assumptions an interpretation as fits their research aims. Motivation for Changing Data Sources The most significant change has been a switch from collecting the raw water rights directly from each state to using the water rights records presented in WestDAAT, a product of the Water Data Exchange (WaDE) Program under the Western States Water Council (WSWC). One of the main reasons for this is that each state of interest is a member of the WSWC, meaning that WaDE is partially funded by these states, as well as many universities. As WestDAAT is also a database with consistent categorization, it has allowed us to spend less time on data collection and quality control and more time on answering research questions. This has included records from water right sources we had previously not known about when creating v1.0 of this database. The only major downside to utilizing the WestDAAT records as our raw data is that further updates are tied to when WestDAAT is updated, as some states update their public water right records daily. However, as our focus is on cumulative water amounts at the regional scale, it is unlikely most records updates would have a significant effect on our results. The structure of WestDAAT led to several important changes to how HarWR is formatted. The most significant change is that WaDE has calculated a field known as `SiteUUID`, which is a unique identifier for the Point of Diversion (POD), or where the water is drawn from. This separate from `AllocationNativeID`, which is the identifier for the allocation of water, or the amount of water associated with the water right. It should be noted that it is possible for a single site to have multiple allocations associated with it and for an allocation to be able to be extracted from multiple sites. The site-allocation structure has allowed us to adapt a more consistent, and hopefully more realistic, approach in organizing the water right records than we had with HarDWR v1.0. This was incredibly helpful as the raw data from many states had multiple water uses within a single field within a single row of their raw data, and it was not always clear if the first water use was the most important, or simply first alphabetically. WestDAAT has already addressed this data quality issue. Furthermore, with v1.0, when there were multiple records with the same water right ID, we selected the largest volume or flow amount and disregarded the rest. As WestDAAT was already a common structure for disparate data formats, we were better able to identify sites with multiple allocations and, perhaps more importantly, allocations with multiple sites. This is particularly helpful when an allocation has sites which cross WMA boundaries, instead of just assigning the full water amount to a single WMA we are now able to divide the amount of water between the number of relevant WMAs. As it is now possible to identify allocations with water used in multiple WMAs, it is no longer practical to store this information within a single column. Instead the stAllocationToWMATab.csv file was created, which is an allocation by WMA matrix containing the percent Place of Use area overlap with each WMA. We then use this percentage to divide the allocation's flow amount between the given WMAs during the cumulation process to hopefully provide more realistic totals of water use in each area. However, not every state provides areas of water use, so like HarDWR v1.0, a hierarchical decision tree was used to assign each allocation to a WMA. First, if a WMA could be identified based on the allocation ID, then that WMA was used; typically, when available, this applied to the entire state and no further steps were needed. Second was the spatial analysis of Place of Use to WMAs. Third was a spatial analysis of the POD locations to WMAs, with the assumption that allocation's POD is within the WMA it should belong to; if an allocation still had multiple WMAs based on its POD locations, then the allocation's flow amount would be divided equally between all WMAs. The fourth, and final, process was to include water allocations which spatially fell outside of the state WMA boundaries. This could be due to several reasons, such as coordinate errors / imprecision in the POD location, imprecision in the WMA boundaries, or rights attached with features, such as a reservoir, which crosses state boundaries. To include these records, we decided for any POD which was within one kilometer of the state's edge would be assigned to the nearest WMA. Other Changes WestDAAT has Allowed In addition to a more nuanced and consistent method of assigning water right's data to WMAs, there are other benefits gained from using the WestDAAT dataset. Among those is a consistent categorization of a water right's legal status. In HarDWR v1.0, legal status was effectively ignored, which led to many valid concerns about the quality of the database related to the amounts of water the rights allowed to be claimed. The main issue was that rights with legal status' such as "application withdrawn", "non-active", or "cancelled" were included within HarDWR v1.0. These, and other water rights status' which were deemed to not be in use have been removed from this version of the database. Another major change has been the addition of the "unspecified water source category. This is water that can come from either surface water or groundwater, or the source of which is unknown. The addition of this source category brings the total number of categories to three. Due to reviewer feedback, we decided to add the acre-foot (AF) column so that the data may be more applicable to a wider audience. We added the ownerClassification column so that the data may be more applicable to a wider audience. File Descriptions The dataset is a series of various files organized by state sub-directories. In addition, each file begins with the state's name, in case the file is separate from its sub-directory for some reason. After the state name is the text which describes the contents of the file. Here is each file described in detail. Note that st is a placeholder for the state's name. stFullRecords_HarmonizedRights.csv: A file of the complete water records for each state. The column headers for each of this type of file are: state - The name of the state to which the allocations belong to. FIPS - The two digit numeric state ID code. siteID - The site location ID for POD locations. A site may have multiple allocations, which are the actual amount of water which can be drawn. In a simplified hypothetical, a farm stead may have an allocation for "irrigation" and an allocation for "domestic" water use, but the water is drawn from the same pumping equipment. It should be noted that many of the site ID appear to have been added by WaDE, and therefore may not be recognized by a given state's water rights database. allocationID - The allocation ID for the water right. For most states this is the water right ID, and what is recommended to use should a right be looked up on a given state's water rights database. The water amounts associated with these IDs tend to be finer scaled than those associated with siteID. It should be noted that some allocations may be extracted from multiple sites, particularly for larger Places of Use. ownerClassification - A classification of the types of owners for water rights. The most common is `Private` which incorporates a wide range of entities. Several classifications would be grouped into a government category, most of which are for the U.S. Federal Government. These allocations could be listed as "Federal", "United States of America", or as the names of any number of federal agencies. The last major grouping of entities is for "Native American"s. priorityDate - The date we use as the water right priority date for our modeling analysis. This is the legal priority date when it is available. However, for some rights, specifically from California and New Mexico, we used a pseudo priority date (e.g. well completion date or start of well drilling date) when a legal priority date was not available. The most questionable dates come from New Mexico, where the only date associated with certain water right records was the date the allocation was recorded in the database. As the allocation record creation tended to be within a few months of the filing of the application of the water right, from manually double checking the water rights, and our analysis focuses on aggregating water rights on the timescale of years, we determined it was acceptable to use such dates to include as many records as possible. primaryBeneficialUse - From the numerous state water use categories, WaDE categorized them into 21 categories WestDAAT. This column is the original WaDE category for the primary water use at the PoD site. allocationBeneficialUse - From the numerous state water use categories, WaDE categorized them into 21 categories for WestDAAT. This column is the original WaDE category

Economics↗

Multiple aspects maintenance ontology-based intelligent maintenance optimization framework for safety-critical systems

Abstract Maintenance optimization is a process for improving the efficiency of maintenance strategies and activities, considering various aspects of the target system and components, such as the probabilities of system failures and the cost of repair and replacement of a failed component. The improvement of maintenance optimization algorithms generally requires information from various data sources. For example, it may require the system risk information derived from risk analysis tools or the residual lifetime of a component from fault prognosis tools. The requirements of data acquisition (DAQ) and aggregation pose new challenges for maintenance management systems (MMSs) that implement and use these maintenance optimization algorithms. This paper proposes a multiple aspects maintenance ontology-based framework to facilitate DAQ from MMSs, online monitoring systems, fault detection and discrimination tools, risk assessment tools, decision-making tools, and component identification tools, and accelerate the implementation and verification of contemporary maintenance optimization models and algorithms. The proposed framework consists of a multi-aspect maintenance ontology with critical information for maintenance optimization and application interfaces for collecting information from various data sources, such as fault prognosis tools, online monitoring tools, risk assessment tools, and decision-making algorithms. In addition, this paper proposes a heuristic method for integrating concepts and properties from other existing ontologies into the proposed framework when the existing ontology is not fully compatible with the ontology under construction. Finally, the paper verifies the proposed ontology framework using a feedwater system designed for nuclear power plants with valves and filters as the components under maintenance.

Diao, Xiaoxu (ORCID:0000000346726352)↗