Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “data guidelines”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Data as a Key Resource in Catalysis: A Community Account

The deployment of artificial intelligence (AI) is transforming the scientific fields central to interdisciplinary catalysis research. By enabling more effective use of data, AI (including simpler machine learning and data science tools) holds great promise for accelerating discoveries. However, progress has so far been modest, largely due to the lack of standardized, machine-readable, and openly shared catalysis data. This perspective, accounting for community insights emerging at conferences, analyses the underlying reasons for these challenges and proposes solutions to a future whereFAIR data management becomes an integral part of research in catalysis. In the short-term, we deem that mandatory FAIR data depositing prior to scientific publications along with consensualized top-down guidelines on data sharing powered by ease-to-use tools can make the necessary step change happen to catalyse data as key resource in our community.

36 - MATERIALS SCIENCE↗

Spinteract: a program to refine magnetic interactions to diffuse scattering data

Magnetic diffuse scattering—the broad magnetic scattering features observed in neutron-diffraction data above a material's magnetic ordering temperature—provides a rich source of information about the material's magnetic Hamiltonian. However, this information has often remained under-utilised due to a lack of available computer software that can fit values of magnetic interaction parameters to such data. Here, an open-source computer program, Spinteract, is presented, which enables straightforward refinement of magnetic interaction parameters to powder and single-crystal magnetic diffuse scattering data. Here, the theory and implementation of this approach are summarised. Examples are presented of refinements to published experimental diffuse-scattering data sets for the canonical antiferromagnet MnO and the highly-frustrated classical spin liquid Gd 3 Ga 5 O 12 . Guidelines for data collection and refinement are outlined, and possible developments of the approach are discussed.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Mechanistic modeling of copper corrosions in data center environments

Air-side economizers are increasingly used to take advantage of “free-cooling” in data centers with the intent of reducing the carbon footprint of buildings. However, they can introduce outdoor pollutants to indoor environment of data centers and cause corrosion damage to the information technology equipment. Here, to evaluate the reliability of information technology equipment under various thermal and air-pollution conditions, a mechanistic model based on multi-ion transport and chemical reactions was developed. The model was used to predict Cu corrosion caused by Cl2-containing pollutant mixtures. It also accounted for the effects of temperature (25 °C and 28 °C), relative humidity (50%, 75%, and 95%), and synergism. It also identified higher air temperature as a corrosion barrier and higher relative humidity as a corrosion accelerator, which agreed well with the experimental results. The average root mean square error of the prediction was 13.7 Å. The model can be used to evaluate the thermal guideline for data centers design and operation when Cl2 is present based on pre-established acceptable risk of corrosion in data centers’ environment.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Demonstration and Evaluation of Explainable and Trustworthy Predictive Technology for Condition-based Maintenance

The domestic nuclear power plant (NPP) fleet has historically relied on labor-intensive and time-consuming predictive maintenance (PdM) programs, thus driving up operation and maintenance (O&M) costs to achieve high-capacity factors. Artificial intelligence (AI) and machine-learning (ML) can help simplify complex problems such as diagnosing equipment degradation to enable more effective decision-making efforts. The benefits of AI will be felt through more efficient plant O&M, improved work processes, and better integration of people and technology. Together, these benefits hold the promise to make nuclear power more sustainable by reducing O&M costs while improving employee engagement. While AI and ML technologies hold significant promise for the nuclear industry, there are challenges or barriers to their adoption. Explainability and trustworthiness of AI are two salient challenges that need to be addressed for wider deployment of these technologies in NPPs. This research focuses specifically on addressing the explainability and trustworthiness of AI technologies to advance the human, technical, and organization (HTO) readiness levels in adopting a risk-informed PdM strategy at commercial NPPs. In addition, this approach can be adapted to enhance the acceptability of AI in other nuclear applications with a few application-specific modifications. The technical approach ensuring wider adoption of AI technologies was developed by Idaho National Laboratory (INL)—in collaboration with Public Service Enterprise Group (PSEG), Nuclear, LLC—by utilizing the circulating water system (CWS) at two PSEG-owned plant sites for demonstration. Focused user studies were performed in collaboration with subject matter experts (SMEs) from PSEG and other nuclear domains to enhance human and organization readiness by building trust in AI-informed technologies. VIsualization for PrEdictive maintenance Recommendation (VIPER)—a Battelle Energy Alliance, LLC, copyrighted software—was developed and expanded to provide a user-centric visualization by incorporating inputs from the collaborating utility, human factors engineering guidelines, and data analysts. The VIPER software enables users, who may be unfamiliar with ML in general, to be interactively engaged by asking technical questions about PdM, work orders, diagnosis results and their confidence levels, the kind of data being used, and the types of ML algorithms employed. This interactive engagement enhances explainability and builds trust. One of the enabling accomplishments was the integration of large language models (LLMs), both text-based and vision-based, in the VIPER software.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

University Data Management Pilot Utilizing the Nuclear Research Data System

Background In 2022, the Office of Science and Technology Policy (OSTP) issued a memo that significantly reshaped the landscape of access to federally funded research. The memo mandated that all taxpayer-funded research be made available to the public without delay upon publication, without an embargo period, superseding the 2013 OSTP public access policy. This public access policy promotes transparency and the democratization of knowledge, ensuring that the fruits of scientific endeavors funded by federal agencies could be immediately accessed and built upon by scientists, educators, students, and the public at large. To implement the requirements of the OSTP guidance and DOE Public Access Plan, the Office of Nuclear Energy (NE) has implemented public access plan guidance and has identified several areas where better data management practices would further expand public access to important nuclear energy related scientific data, reports, and other technical products. Significant NE supported efforts are already underway for data management and public access to important nuclear energy related data.1 2 To address gaps in data management practices, and improve retention and accessibility of data, NE is actively exploring enhanced data management options utilizing its high-performance computing resources administered by its Nuclear Scientific User Facility Program. A newly piloted system, the Nuclear Research Data System (NRDS) acts as a portal for data collection and dissemination. Nuclear Energy University Program Research and Development Portfolio According to Web of Science, NEUP has produced 2,345 journal publication that have been cited more than 61,000 times3 and countless conference proceedings. These publications are publicly available through OSTI.gov and in the open literature. Additional scientific and technical products including project milestones that are not publications and NEUP project final reports are vetted through OSTI.gov and released once reviewed and approved by DOE. Since 2009, NEUP has awarded close to 1,000 different R&D projects in technical areas across the NE research programs. As of June 2023, 512 NEUP reports are publicly available on OSTI. The underlying data for projects is still held at universities, and data transfer, co-location, and dissemination has not occurred in a systematic way. NEUP data is currently accessible through myriad university-based data repositories, or through direct requests to PIs. The program identified this patchwork of repositories, or often lack of publicly available data, as a significant barrier to an organized, accessible, and comprehensive solution to sharing data with the larger nuclear energy community. Approach The goal of this pilot project is to establish a pathway to a consolidated long-term repository for NEUP project data. To accomplish this goal, the pilot strives to accomplish the following objectives: Establish data collection standards, including a standard set of required supplementary information to contextualize and support raw data files. Work with the HPC group collect and upload information and to modify the NRDS system, as needed, to support a standardized approach. Resolve potential barriers to successful roll out of an expanded data collection strategy, including modifying data management plan guidelines and establishing a document and data release process that accounts for potential intellectual property and/or export control concerns. Results Overall, the pilot was successful in collecting 8,982 raw and processes data files, 220 reports, 56 calibration files, and 5,931 other supplementary documents. Supplementary documents included experimental plans, methods, journal publications and conference proceedings, milestone reports, and final reports. Figure 2 shows the number of data sets and supplementary project information provided by each project. Projects has significantly different input, depending on experimental data produced and completeness of the datasets provided.

Data collection↗

ASME Section III, Division 5, High Temperature Reactors

Section III, Division 5 stands ready to support near-term deployment of advanced reactors. Progress on improving design rules, extending design lifetimes, adding more materials, and adding advanced component fabrication methods. Lessons-learned from Alloy 617 Code Case effort have streamlined the balloting workflow for Class A material code cases. After requester submitted material design parameters and supporting data package to ASME, Division 5 could turn around a material code case in about three Code Week cycles (less than a year). Data requirements for new materials are described in Division 5, “Nonmandatory Appendix HBB-Y, Guidelines For Design Data Needs For New Materials”.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Higher than expected N 2 O emissions from soybean crops in the Pampas Region of Argentina: Estimates from DayCent simulations and field measurements

In developing countries, agriculture generally represents a large fraction of GHG emissions reported in National Inventories, and emissions are typically estimated using Tier 1 IPCC guidelines. However, field data and locally adapted simulation models can improve the accuracy of IPCC estimations. In this report we aimed to quantify anthropogenic N2O emissions from croplands of Argentina through field measurements, model simulations and IPCC guidelines. Here we measured N 2 O emissions and their controlling factors in 62 plots of the Pampas Region with corn, soybean and wheat/soybean crops and in unmanaged grasslands. We accounted for gross emissions from crops and background emissions from unmanaged grasslands to calculate net anthropogenic emissions from crops as the difference between them. We calibrated and evaluated the DayCent model and then simulated different weather and management scenarios. Finally, we applied IPCC guidelines to estimate anthropogenic N 2 O emissions at the same plots. The DayCent model accurately simulated annual N 2 O emission for all crops as compared to measured data (RMSE = 1.4 g N ha -1 day -1 ). Measured and simulated emissions in soybean crops were higher than in corn and wheat/soybean crops. Gross N 2 O emissions ranged from 1.4 to 5.1 kg N ha -1 yr -1 for current environmental (soil and weather) and management (crops and fertilizer doses) conditions. Background emissions ranged between 1.1 and 1.3 kg N ha -1 yr -1 , and therefore net anthropogenic emissions ranged from 0.3 to 4.0 kg N ha-1 yr -1 . IPCC Tier 1 emission factors underestimated N2O releases from soybean, that were on average 4.87 times greater when estimated with DayCent and observations (0.53 vs 2.47 and 2.69 kg N ha -1 yr -1 , respectively). On the contrary, IPCC estimates for corn and wheat/soybean crops were similar to modeled and measured values. Our results suggest that N 2 O emissions from the vast 15 million ha of soybean croplands in the Pampas Region may be substantially underestimated.

59 BASIC BIOLOGICAL SCIENCES↗

Detection and Perception of Sound by Eagles and Surrogate Raptors

One overarching objective of this program of study was the accumulation of objective, scientifically valid information relating to auditory performance of bald and golden eagles that may be used to guide the development of acoustic alerting/deterrence technologies intended to discourage encroachment into wind energy air spaces. To that end, analyses aimed at the characterization of sensitivity to sound in bald and golden eagles, along with findings in the supra-threshold, dynamic frequency spaces related to response latencies and amplitudes, leads us to conclude that bald, and golden eagles navigate the same basic working auditory space, as in other known and thus far characterized members of the diurnal raptor family. Specifically, bald and golden eagles, along with other raptor species within the group, operate in an auditory space characterized by a frequency band at least four octaves wide and centered on 2 kHz, with an upper frequency limit between 6 and 10 kHz at 80 dB SPL and a lower frequency limit that almost certainly extends below 0.2 kHz. Consequently, we recommend that signal designers use these data as a guideline in efforts to design effective and efficient acoustic alerting/deterrent systems. It is important to note that signal energy broadcast outside of this frequency band at moderate levels will not contribute to the efficacy of a deterrent but will add an unnecessary fraction to the overall acoustic pollution budget. The importance of this consideration is heightened by contemporaneous concerns related to the transmission of noise broadcast by wind energy farms. In addition, based on analyses of data acquired from red-tailed hawks using the same experimental paradigm and data acquisition system, we conclude that auditory function in the red-tailed hawk is sufficiently like that observed in bald and golden eagles to permit its use as a surrogate species. Response waveforms, threshold-frequency curves, and input-output characteristics match those of eagles closely. It should be noted however, that differences in sensitivity and slightly extended high-frequency limits of hearing should be taken into account when extrapolating findings from one species to the others. Although the inclusion of behavioral tests of red-tailed hawks to acoustic stimuli was beyond the scope of this investigation, future efforts to assess response parameters like signal-type preference and habituation rate will further elucidate their suitability to serve as eagle surrogates in behavioral studies; nonetheless, the species in question are well matched with respect to basic auditory performance. A second essential objective of this program of study was the acoustic characterization of a subset of calls comprising the vocal repertoires of bald and golden eagles that may be used to supplement auditory performance findings in the effort to guide the development of acoustic alerting signals. With regard to that objective, the vocal repertoires of both bald and golden eagle species are rich and varied. While similar in spectrographic structure, distinctive differences are also clear. Generally, golden eagles produce some calls with shorter durations, and similar “sounding” calls exhibit distinctively different spectrographic patterns than those of bald eagles. Both species produce calls that contain a wide variety of nonlinear elements that operate to enhance the rich and varied nature of commonly observed vocal products. Comparison of the average power spectra of commonly observed bald and golden eagle calls with threshold-frequency curves leads to the conclusion that call energies fall within the frequency bounds of hearing. Further, the acoustic energy of calls considered in this report tend to fall into overlapping, but different frequency ranges of the acoustic sensitivity curve. This condition may encourage signal designers to vary the frequency content of acoustic deterrence signals in the field. Finally, preliminary observations relating to the tendencies and proclivities of bald eagles to attend to the acoustic landscape lead to the conclusion that eagles monitor their immediate sound environment assiduously. Individuals respond to a variety of natural and synthetic sound signals reliably and, perhaps most relevant in the context of the engineering of acoustic alerting/deterrence technologies, habituation to most sounds considered in this effort was minimal. These preliminary results, while calling for extended behavioral testing, are promising and set the stage for the exportation of behavioral studies into real world scenarios.

17 WIND ENERGY↗

A standards perspective on genomic data reusability and reproducibility

Genomic and metagenomic sequence data provides an unprecedented ability to re-examine findings, offering a transformative potential for advancing research, developing computational tools, enhancing clinical applications, and fostering scientific collaboration. However, effective and ethical reuse of genomics data is hampered by numerous technical and social challenges. The International Microbiome and Multi’Omics Standards Alliance (IMMSA, https://www.microbialstandards.org/) and the Genomic Standards Consortium (GSC, https://gensc.org) hosted a 5-part seminar series “A Year of Data Reuse” in 2024 to explore challenges and opportunities of data reuse and reproducibility across disparate domains of the genomic sciences. Addressing these challenges will require a multifaceted approach, including common metadata reporting, clear communication, standardized protocols, improved data management infrastructure, ethical guidelines, and collaborative policies that prioritize transparency and accessibility. We offer strategies to enable responsible and technically feasible data reuse, recognition of data reproducibility challenges, and emphasizing the importance of cross-disciplinary efforts in the pursuit of open science and data-driven innovation.

59 BASIC BIOLOGICAL SCIENCES↗

ESS-DIVE Reporting Format for Comma-separated Values (CSV) File Structure

The ESS-DIVE reporting format for Comma-separated Values (CSV) file structure is based on a combination of existing guidelines and recommendations including some found within the Earth Science Community with valuable input from the Environmental Systems Science (ESS) Community. The CSV reporting format is designed to promote interoperability and machine-readability of CSV data files while also facilitating the collection of some file-level metadata content. Tabular data in the form of rows and columns should be archived in its simplest form, and we recommend submitting these tabular data following the ESS-DIVE reporting format for generic comma-separated values (CSV) text format files. In general, the CSV file format is more likely accessible by future systems when compared to a proprietary format and CSV files are preferred because this format is easier to exchange between different programs increasing the interoperability of a data file. By defining the reporting format and providing guidelines for how to structure CSV files and some field content within, this can increase the machine-readability of the data file for extracting, compiling, and comparing the data across files and systems.Data package files are in .csv, .png, and .md. Open the .csv with e.g. Microsoft Excel, LibreOffice, or Google Sheets. Open the .md files by downloading and using a text editor (e.g., notepad or TextEdit). Open the .png in e.g. a web browser, photo viewer/editor, or Google Drive.

54 ENVIRONMENTAL SCIENCES↗

Ground-based_transient_electromagnetic_data_and_resistivity_models_for_SubTER_BelleCreek_2017-2018

Ground-based transient electromagnetic (TEM) data were acquired in selected locations around Belle Creek, Montana to define the resistivity structure of the near surface. Data were acquired using the ABEM WalkTEM system (Guideline Geo Ab, Sundbyberg, Sweden). TEM data were processed and numerically inverted to derive one-dimensional resistivity structure using SPIA (Aarhus Geosoftware, Aarhus, Denmark). This release includes raw and processed TEM data and resistivity models for each data location using the following structure: Ground-based_transient_electromagnetic_data_and_resistivity_models_for_SubTER_BelleCreek_2017-2018.xml SBC_TEM_2019_SoundingLocationsModels.gdb.zip - file geodatabase with all models and plots attached. This is a mirror of the contents of the model directory. Ground-based_transient_electromagnetic_data_and_resistivity_models_for_SubTER_BelleCreek_2017-2018.zip Data SBC_USF - raw data USF_format_description.pdf - manual describing the USF file format SBC_SoundingID-USF_lookup.csv - table describing the USF file naming convention and channel mapping *.usf files - raw, unprocessed data files SBC_TEM - processed data AarhusInv_manual.pdf - manual describing the TEM file format *.tem files - processed data files Models SBC_Model_Plots - plots of the data and the models for each sounding location TEM_model_data_dictionary.csv - data dictionary describing tabulated models SBC_TEM_models.csv - tabulated models

geophysical surveying↗

Catalytic Upgrading of Pyrolysis Products for the Production of Sustainable Aviation Fuel

The objective of this project is advance the state-of-technology for a catalytic fast pyrolysis (CFP) + hydrotreating (HT) process to produce sustainable aviation fuel and other biogenic products. Our approach focuses on performing integrated experiments using realistic biomass feedstocks and non-noble metal technical catalyst formulations. CFP is performed in an ex-situ configuration using a fluidized bed reactor without co-fed hydrogen. Research advancements over the past two years include establishing benchmark yield structures and compositional data for each step of the biomass-to-SAF process, demonstrating the ability to produce a cycloalkane-rich SAF product that meets key ASTM 4054 guidelines, generating benchmark characterization data for technical catalyst formulations with an emphasis on determining the unique composition and combustion properties of biogenic coke, and establishing bio-oil critical material attributes to mitigate the risk of plugging during down-stream hydroprocessing. Other impacts from this project include generation of broadly enabling scientific knowledge (12 publications/12 presentations since 2021), engagement with industry partners (Johnson Matthey, ExxonMobil, Phillips 66), and identification of a promising pathway to market that addresses emerging demands for biogenic refinery feedstocks.

biomass↗

Multi-Variable Parametric Analysis of Prototype Building Energy Performance Using Current and Future Weather Scenarios For Data-Driven Market Transformation Support

This project aimed to develop a public building simulation data set that may be used to inform building code development and guidelines for building innovation. The data set consists of several common building types and many representative locations across the United States. A parametric design of building properties was developed to create a range of building energy models that represent common building design decisions with a particular focus on fenestration options. The US Department of Energy prototype building energy models were altered according to a parametric building design and simulated using both current weather data and future weather estimates derived from global climate models. The resulting data set allows for pertinent exploration of building design parameters, including fenestration, within different environments across the United States in the broader context of climate change.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Best Practices Guide for Energy-Efficient Data Center Design

This guide provides an overview of best practices for energy-efficient data center design which spans the categories of information technology (IT) systems and their environmental conditions, data center air management, cooling and electrical systems, and heat recovery. IT system energy efficiency and environmental conditions are presented first because measures taken in these areas have a cascading effect of secondary energy savings for the mechanical and electrical systems. This guide concludes with a section on metrics and benchmarking values by which a data center and its systems energy efficiency can be evaluated. No design guide can offer “the most energy-efficient” data center design but the guidelines that follow offer suggestions that provide efficiency benefits for a wide variety of data center scenarios.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Energy Efficient Data Centers

Energy efficient data center strategies include thermal/environmental guidelines, air management, free cooling, and liquid-based cooling. A case study highlighting the data center cooling system at the National Renewable Energy Laboratory Energy Systems Integration Facility illustrates many of these principles.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Energy Efficient Data Centers

Energy efficient data center strategies include thermal/environmental guidelines, air management, free cooling, and liquid-based cooling. A case study highlighting the data center cooling system at the National Renewable Energy Laboratory Energy Systems Integration Facility illustrates many of these principles.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Quality Control Methods for Advanced Metering Infrastructure Data

While urban-scale building energy modeling is becoming increasingly common, it currently lacks standards, guidelines, or empirical validation against measured data. Empirical validation necessary to enable best practices is becoming increasingly tractable. The growing prevalence of advanced metering infrastructure has led to significant data regarding the energy consumption within individual buildings, but is something utilities and countries are still struggling to analyze and use wisely. In partnership with the Electric Power Board of Chattanooga, Tennessee, a crude OpenStudio/EnergyPlus model of over 178,000 buildings has been created and used to compare simulated energy against actual, 15-min, whole-building electrical consumption of each building. In this study, classifying building type is treated as a use case for quantifying performance associated with smart meter data. This article attempts to provide guidance for working with advanced metering infrastructure for buildings related to: quality control, pathological data classifications, statistical metrics on performance, a methodology for classifying building types, and assess accuracy. Advanced metering infrastructure was used to collect whole-building electricity consumption for 178,333 buildings, define equations for common data issues (missing values, zeros, and spiking), propose a new method for assigning building type, and empirically validate gaps between real buildings and existing prototypes using industry-standard accuracy metrics.

24 POWER TRANSMISSION AND DISTRIBUTION↗

HarDWR - Raw Water Rights Records

A dataset within the Harmonized Database of Western U.S. Water Rights (HarDWR). For a detailed description of the database, please see the meta-record v2.0. Changelog v2.0 - Switched source data from collecting records from each state independently to using the WestDAAT dataset v1.0 - Initial public release Description In order to hold a water right in the western United States, an entity, (e.g., an individual, corporation, municipality, sovereign government, or non-profit) must register a physical document with the state's water regulatory agency. State water agencies each maintain their own database containing all registered water right documents within the state, along with relevant metadata such as the point of diversion and place of use of the water. All western U.S. states have digitized their individual water rights databases, as well as geospatial data defining the areas in which water rights are managed. Each state maintains and provides their own water rights data in accordance with individual state regulations and standards. In addition, while all states make their water rights publicly available, each provides their records in unique formats, meaning that file types, field availability, and terms vary from state to state. This leads to additional challenges to managing resources which cross state lines, or conducting consistent multi-state water analyses. For the first version of HarDWR, we collected the water rights databases from 11 Western States of the United States. In order to preform regional analyses with the collected data, the raw records had to be harmonized into one single format. The Water Data Exchange (WaDE) is a program dedicated to the sharing of water-related data for the Western U.S. in a singular consistent format. Created by the Western States Water Council (WSWC) to facilitate the collection and dissemination of water data among WSWC's member states and the public, WaDE provides an important service for those interested in water resource planning and management in their focus region. Of the services which WaDE provides, the one of the most interesting is the WestDAAT dataset, which is a collection of water rights data provided by the 18 WSWC member states that have been standardized into a single format, much like we had done on a more limited scale with HarDWR v1. For this version of HarDWR we decided to use WestDAAT, specifically a snapshot created in Feburary 2024, as our water rights source data. A full explanation of the benefits gained from this switch can be found in the description of the updated Harmonized Water Rights Records v2.0, but in short it has allowed us to focus more of our efforts on answering research questions and gaining a more realistic understanding of how water rights are allocated. For more information on how the data for WestDAAT was collected, please see the WaDE data summary. Terms of Use While WaDE works directly with the state agencies to collect and standardize the water rights records, the ultimate authority for the water rights data remains the individual states. Each state, and their respective water right authorities, have made their water right records available for non-commercial reference uses. In addition, the states make no guarantees as to the completeness, accuracy, or timeliness of their respective databases, let alone the modifications which we, the authors of this paper, have made to the collected records. None of the states should be held liable for using this data outside of its intended use. As several of the states update their water rights databases daily, the information provided here is not the latest possible, and should not be used for legal purposes. WestDAAT itself has irregular updates. Additional questions about the data the source states provided should be directed to the respective state agencies (see methods.csv and organization.csv files described below). In addition, although data was presented here was not collected directly from the states, several states requested specifically worked disclaimers when sharing their data. These disclaimers are included here as an acknowledgement from where the water rights data is primarily sourced. Colorado: "The data made available here has been modified for use from its original source, which is the State of Colorado. THE STATE OF COLORADO MAKES NO REPRESENTATIONS OR WARRANTY AS TO THE COMPLETENESS, ACCURACY, TIMELINESS, OR CONTENT OF ANY DATA MADE AVAILABLE THROUGH THIS SITE. THE STATE OF COLORADO EXPRESSLY DISCLAIMS ALL WARRANTIES, WHETHER EXPRESS OR IMPLIED, INCLUDING ANY IMPLIED WARRANTIES OF MERCHANTABILITY, OR FITNESS FOR A PARTICULAR PURPOSE. The data is subject to change as modifications and updates are complete. It is understood that the information contained in the Web feed is being used at one's own risk." Montana: "The Montana State Library provides this product/service for informational purposes only. The Library did not produce it for, nor is it suitable for legal, engineering, or surveying purposes. Consumers of this information should review or consult the primary data and information sources to ascertain the viability of the information for their purposes. The Library provides these data in good faith but does not represent or warrant its accuracy, adequacy, or completeness. In no event shall the Library be liable for any incorrect results or analysis; any direct, indirect, special, or consequential damages to any party; or any lost profits arising out of or in connection with the use or the inability to use the data or the services provided. The Library makes these data and services available as a convenience to the public, and for no other purpose. The Library reserves the right to change or revise published data and/or services at any time." Oregon: "This product is for informational purposes and may not have been prepared for, or be suitable for legal, engineering, or surveying purposes. Users of this information should review or consult the primary data and information sources to ascertain the usability of the information." File Descriptions The unmodified February, 2024 WestDAAT snapshot is composed of nine files. Below is a brief description of each file, as well as how they were utilized for HarDWR. WaDEDataDictionaryTerms.xlsx: As the file's name implies, this is a data dictionary for all of the below named files. This file describes the column names for each of the following files, with the exception of citation.txt which does not have any columns. The descriptions for each file are divided by tab,with the same name as their associated file, within this document. allocationamount.csv: The "main" file of the group, it contains the water right records for each state. Of particular note, each water right is broken down into one or more water allocations. Allocations may be withdrawn from one or more locations, or even multiple allocations associated with a particular location. This is a more subtle and realistic representation of how water is used than what was available in the first version of HarDWR. For the records from some states, this can mean that multiple allocations listed under a single right will appear as rows within this file. citation.txt: A combination of contact information for WaDE personnel, disclaimer about how the data should be used, and guidelines for citing WestDAAT. methods.csv: A file describing the source and method by which WaDE collected water rights data from each state. organization.csv: A file listing the water rights authoritative agencies for each state. sites.csv: This file provides the geographic, and other descriptors, of the physical location of allocations, called 'sites'. To reiterate, it is possible for one allocation to be associated with multiple sites, as well as one site to be associated with multiple allocations. The two descriptors which we were most interested in where the site's coordinates, as well as whether the site was classified as a Point of Diversion (POD) or a Place of Use (POU). As a general rule, PODs are geographic points, while POUs are areas typically represented as property boundaries or irregularly shaped polygons. sites_pouGeometry.csv: For those allocations with a POU site, this file contains the defining points for the associated polygons. variables.csv: A file describing the units in which an allocation's water amount is reported within WestDAAT. This information is essentially a repeat of the 'AllocationFlow_CFS' and 'AllocationVolume_AF' columns within allocationamount.csv, at least for our purposes. watersources: This file describes the source of water from which each site extracts from. For our purposes, this table was used to determine whether the water came from Surface Water, Groundwater, or Unspecified Water.

Lisk, Matthew↗