Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Anonymization”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

Bay Area Regional Energy: Network Integrated Commercial Retrofits (BRICR) Project. Final Report

The BRICR project applied large-scale building energy modeling concepts with the aim of reducing the cost of energy efficiency targeting, design, and project development, and measurement of energy savings for energy efficiency programs implemented by local governments that serve small and medium commercial buildings (SMB). The project leveraged the services and resources of existing local government energy programs serving disadvantaged and hard-to-reach SMB customers. In contrast to programs run by utilities, local government programs generally do not have direct access to energy billing records for an entire class of customers in a geographic area, which prior research demonstrated useful for large-scale building energy model baseline development and calibration. , However, local governments are rich in public records that offer important clues about physical attributes and uses that, along with behavior, determine energy use. Relying only on public records, BRICR demonstrated development of credible baseline energy models for 3,792 office, retail, and hotel buildings. Publicly disclosed annual energy use data from a local energy benchmarking program and anonymized data from the Building Performance Database, the nation’s largest dataset about energy-related characteristics of buildings, were utilized to validate and calibrate energy models via an innovative method comparing distributions of energy intensity by fuel type for portfolios of buildings of similar size, vintage, and use. Portfolio calibration does not provide certainty that an energy model fits an individual building; the method is useful when billing data is not accessible – a common situation for researchers, energy service providers and ESCOs, local governments, and any party other than a utility. A software component was developed, the BRICR gem, which automates simulation when relevant data is added or edited by the user to a file saved in the standardized BuildingSync XML schema for energy audit data. The component was demonstrated as a simplified means to generate a mass of energy models corresponding to public records containing basic attributes such as building scale, location, use, year built, and aspect ratio in combination with building energy code prototype data corresponding to use and vintage. The component was also demonstrated as a simplified means to automate energy simulation when attributes are revised; the intention was to enable iterative improvement of the baseline model and energy savings estimates for common energy conservation measures as users revise relevant attributes based on their observations. In the context of institutional change and uncertainty for the participating local government energy programs, 13 whole building retrofits were completed. Impacts were measured by applying the CalTRACK2.0 methods to standardize measurement of normalized metered energy consumption. The GRIDMeter methods of stratified sampling and individual load shape analysis were applied to adjust for impacts of the effect of COVID-19 on retrofitted buildings in the context of all local buildings of similar size and use. Excluding impacts of the pandemic, retrofitted buildings demonstrated between 1.6% and 25.1% reduction in energy use. The project contributed use cases and feedback that helped inform evolution of the software tools and data formats that were combined for the first time in the BRICR project.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Assessment of BQ-9000 Biodiesel Properties for 2021

This is the fifth in a series of reports documenting the quality of biodiesel from U.S. and Canadian-based producers that participate in the BQ-9000 program, the biodiesel industry voluntary quality assurance program. Participants agreed to provide monthly data on critical quality parameters for calendar year 2021. The quality data was provided to a team of experts, who removed any identifying company information and provided anonymized and randomized data to the National Renewable Energy Laboratory (NREL) for statistical analysis. The critical quality parameters analyzed were: sodium and potassium (Na+K); calcium and magnesium (Ca+Mg); phosphorus (P); flash point and alcohol control; water and sediment; cloud point; acid number; free and total glycerin; monoglycerides; sulfur; oxidation stability; and cold soak filterability test (CSFT). The data was not weighted for production volume.

09 BIOMASS FUELS↗

Development and Validation of Algorithms That Analyze Communicating Thermostat Data to Identify Enclosure Retrofit Opportunities

Annual energy savings of up to $\$ 4$ to $\$ 5$ billion could be achieved nationwide through basic insulation and heating system retrofits of existing homes. However, current utility energy efficiency programs are costly and challenging to scale. Customer acquisition occurs primarily through energy bill mailers, mass media, and online advertising that lack specificity about home-specific retrofit opportunities, expected energy savings, and cost-effectiveness. Specific retrofit opportunities are identified via on-site home energy assessments (HEAs) that are inconvenient to homeowners, expensive, and of variable accuracy. We developed computational algorithms that automatically analyze communicating thermostat (CT) heating data that could be used to increase the customer uptake of insulation and air sealing energy conservation measures (ECMs) by identifying homes with the most significant retrofit opportunities, estimating post-retrofit energy savings, and formulating home-specific outreach. The algorithms are based on an extended second-order grey-box model that characterizes a building’s thermal response using lumped elements, coupled with an empirical model of infiltration that accounts for both wind and stack effects. The basic parameters of the model correspond to actual physical parameters of the home, i.e., the home’s overall R-value of and the building envelope ACH50. Unlike the conventional approach, which estimates model parameters based on the best fit to the observed time-dependent room temperature, our approach derives correlations between the daily heating system runtime and temperature difference (indoor-outdoor) that are more robust to data quality issues in real-world applications. We also used HEA data for algorithm development and validation. With the help of our utility partners, Eversource and National Grid, we obtained data sets for hundreds of Massachusetts homes. For each home, these data sets included three sets of information anonymized by the utility: (1) CT data (HVAC runtime, room temperature, and, for some vendors, outdoor temperature and wind speed) collected by the CT vendor (one of three) over a heating season, (2) HEA report performed by the HEA vendor (same vendor for all homes), (3) Monthly utility gas bills coincident with the CT data (3 to 24 per home, depending on availability). For some homes, we also obtained blower-door test results. Initially, we applied the algorithms developed to homes with a single CT and then extended them to homes with two CTs by using an equivalent home approach. Finally, we developed algorithms for prediction of energy savings and a methodology of comparing our predictions with those generated by HEAs. The main technical results indicate that we can reliably identify homes with insulation and/or air sealing retrofit opportunities and provide accurate savings predictions. Our hypothesis is that the algorithms could be applied to utility energy efficiency programs to identify homes that could realize significant energy savings from insulation and/or air sealing retrofits. This information could then be used to reach out to those homes with highly customized outreach, thereby delivering increased program energy savings and cost-effectiveness. This would: Significantly increase the uptake rate of on-site HEAs, and Significantly increase the fraction of HEAs resulting in ECM implementation. To test these hypotheses, we designed and conducted a randomized controlled trial (RCT). The RCT results suggest that personal messaging leads to a two- to five-fold increase in the HEA uptake rate.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Reducing Uncertainty of Fielded Photovoltaic Performance (Final Technical Report)

Improved analysis and reporting of photovoltaic (PV) field performance increases the certainty of owners and financiers that systems will perform as expected. Advanced module technologies (e.g., PERC, HJT, and bifacial) introduce new degradation mechanisms and performance characteristics. The FY19-21 Reducing Uncertainty project leveraged data from the ever-increasing PV fleet to develop models and understanding of the field performance of existing and new technologies. Specifically, we accomplished: report on field performance and degradation rates for high-efficiency silicon (HJT, PERC, IBC) and more conventional technologies; developed automated analysis techniques to quantify system performance (performance ratio, energy yield) and production shortfalls (soiling, degradation, availability); refined the RdTools software toolkit to bring standard, validated analysis techniques to bear on third-party data; analyzed and reported on large datasets including Treasury data and Lawrence Berkeley National Laboratory's Utility-Scale dataset to expand the high-quality degradation-rate histogram published previously; worked with industry partners and the DuraMAT data hub to enable private parties to share and aggregate PV production data anonymously, leveraging cloud-based data analysis infrastructure and publishing on US fleet-scale performance comprising over 7GW of operating systems. (https://www.nrel.gov/pv/fleet-performance-data-initiative.html). Through our industry collaborations we have engaged in NDA-covered data transfer with twelve PV fleet owners as of January 2022, with more agreements in negotiation. Our scalable cloud-based time series database contains over 30 billion rows (20TB) of PV time series data, representing over 1700 commercial and utility-scale systems, and over 7.2 GW of DC capacity (Fig 1). Initial field performance results have been distributed in several public reports. Because our fleet composition and data quality methods are continually improving, annual updates to these results are published to our PV Fleet webpage [ https://www.nrel.gov/pv/fleet-performance-data-initiative.html ] and DuraMAT data hub [DOI: 10.21948/1842958]. Another existing dissemination channel used for observed soiling losses is a map we maintain for soiling losses. Additional products developed include a report detailing fleet-wide performance index, availability, startup loss and snow loss factors, a detailed report on the 1603 grant dataset comprising over 100,000 PV systems with failure and performance details and a utility-scale report coauthored with LBNL on 31 GW of system performance.

14 SOLAR ENERGY↗

Analysis of Dust Samples Collected from a Near-Marine East Coast ISFSI Site ("Site C")

In June of 2022, dust samples were collected from the surface of an in-service spent nuclear fuel dry storage canister during an inspection at an Independent Spent Fuel Storage Installation. The site is anonymous but is a near-marine or brackish water east coast location referred to here as "Site C". The purpose of the sampling was to assess the composition and abundance of the soluble salts present on the canister surface, information that provides a metric for potential corrosion risks. Following collection, the samples were delivered to Sandia National Laboratories for analysis. At Sandia, the soluble salts were leached from the dust and quantified by ion chromatography. In addition, subsamples of the dust were taken for scanning electron microscopy to determine the particle sizes, morphology, and mineralogy of the dust and salts. The results of those analyses are presented in this report.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

SARS-CoV-2 transmission control in California correctional facilities

LBNL provided technical assistance to the CalPROTECT project, which is a collaboration between the University of California, Berkeley and the University of California, San Francisco. The CalPORTECT team provided support and guidance to the California Department of Corrections and Rehabilitation (CDCR) and the California Correctional Health Care Services (CCHCS) in their response to the COVID-19 pandemic. The CalPROTECT team provided recommendations to mitigate the risk posed to the residents and staff who live and work in CDCR prisons through considering scientific knowledge of the virus, analyzing anonymized resident and staff level data from CCHCS, evaluation of statewide and institution-level policies, and data collection during site visits to CDCR prisons during the pandemic.

59 BASIC BIOLOGICAL SCIENCES↗

Assessment of BQ-9000 Biodiesel Properties for 2022

This is the sixth in a series of reports documenting the quality of biodiesel from U.S. and Canadian-based producers that participate in the BQ-9000 program, the biodiesel industry voluntary quality assurance program. Participants agreed to provide monthly data on critical quality parameters for calendar year 2022. The quality data was provided to a team of experts, who removed any identifying company information and provided anonymized data to the National Renewable Energy Laboratory (NREL) for statistical analysis.

09 BIOMASS FUELS↗

Availability and Performance Loss Factors for U.S. PV Fleet Systems

In the PV Fleet Performance Data Initiative, we partner with photovoltaic (PV) fleet owners to collect time-series PV production data and publish aggregated, anonymized results. This report is an update of our previous publications, specifically a FY 2021 performance index publication and a FY 2022 fleet degradation analysis. In this analysis, we have increased our data participants and system totals by around 10% to 8.5 GW and 24,000 separate inverter data channels. Four major analysis topics are considered in this report: Performance Index (PI) trends, PV system availability, soiling losses, and PV system degradation. Performance Index and inverter availability are assessed on a larger set of data from our FY 2021 report: 1,128 systems compared with 200 systems from before. The increased number of systems is due to an improved data quality methodology, as well as introducing new systems to the analysis. Overall results are similar to previously published values - overall inverter availability is low in the first six months of system performance before reaching steady-state by the end of the first year. Excluding this six-month startup period, system-level aggregated data shows a median (P50) system availability of 0.99 and a lower 10th percentile (P90) value of 0.95 (Figure ES-1). A dependence on system size is also demonstrated, with worse inverter availability results for larger PV systems. Causes of this effect are under investigation, but may be impacted by inverter size, which also show lower availability for larger inverter sizes. This report also investigates PI, correcting for degradation, soiling, snow, and availability. Following these corrections, the median system PI over its entire lifetime is 0.95. PI values reported here are approximately 3% lower than what we presented in our previous FY 2021 report. Soiling loss is assessed in a comprehensive way for the first time in this report. Results are presented using the COmbined Degradation and Soiling (CODS) method, as implemented in RdTools (v3.0.0a4). Soiling values are presented for 255 systems, which indicated irradiance-weighted soiling loss greater than 1%. The values have been published in an updated NREL soiling map at nrel.gov/pv/soiling.html. Finally, we investigated system degradation using three different data analysis techniques: conventional RdTools (year-on-year (YOY)), CODS, and Performance Loss Rate (PLR) analysis. Overall degradation results are consistent with our previous publications. Rerunning conventional RdTools on our updated fleet shows that some data partners have systematically fallen below the median system degradation rate (change over time) of -0.75 %/year. A comparison with PLR analysis, which looks at change in annual PI over time, shows that median system degradation is consistent with -0.5% to -0.75% per year change. However, at the P90 value, system degradation is substantially faster. These two results are consistent and indicate that resulting degradation statistics depend to a great degree on the population of PV systems making up the analysis cohort and whether soiling impacts the systems. The use of CODS for degradation analysis provides a different method for degradation assessment, which explicitly excludes the impact of recoverable soiling on degradation analysis. Excluding soiling effects yields an annual system degradation around -0.5% per year on average. This indicates that a portion of system performance loss may be attributed to periodic soiling that is not fully recovered. This report provides PV system owners/operators with background and methods to analyze PV system performance, give guidance for expected cohort performance, and performance loss values for use in pro-forma financial models, which guide new-build system design and bankability reports.

14 SOLAR ENERGY↗

PV System Availability from Commercial and Utility-Scale Systems [Slides]

In the PV Fleet Performance Data Initiative, we partner with photovoltaic (PV) fleet owners to collect time-series PV production data, and publish aggregated, anonymized results. An assessment of system availability is conducted on 1128 systems which passed our data quality checks, and include cumulative energy meter data. Overall inverter availability is low in the first 6 months of system performance before reaching steady-state by the end of the first year. System-level aggregated data shows a median (P50) system availability of 0.99, and a lower P90 value of 0.95. A dependence on system size is also identified, with worse inverter availability results for larger PV systems. Potential causes of this effect are under investigation.

14 SOLAR ENERGY↗

Assessment of BQ-9000 Biodiesel Properties for 2023

This is the seventh in a series of reports documenting the quality of biodiesel from U.S.- and Canadian-based producers that participate in the BQ-9000 program, the biodiesel industry’s voluntary quality assurance program. Participants provided monthly data on critical quality parameters for calendar year 2023 with quality data provided to a team of experts, who removed any identifying company information and provided anonymized data to the National Renewable Energy Laboratory (NREL) for statistical analysis. New for 2023, data on kinematic viscosity, sulfated ash, distillation temperature, carbon residue, and cetane were collected, as well as individual levels of sodium, potassium, calcium, and magnesium. Critical quality parameters analyzed are listed in Table ES-1 with descriptive statistics.

09 BIOMASS FUELS↗

Evaluating a Commercial Dynamic Line Rating Software with the National PMU Dataset

To accelerate the development of data-driven applications for power systems, the Department of Energy (DOE) supported the collection and curation of a synchrophasor dataset spanning two years of observations from transmission utilities across the US. This National PMU Dataset (NPDS) was anonymized and distributed to awardees of a DOE research grant under nondisclosure agreements (NDAs) but has also been retained at PNNL to enable further research. Agreements with data contributors prevent the data from being shared outside the organization. However, establishing a blind research validation methodology is envisioned to maximize the value proposition of the NPDS. In this validation strategy, researchers may share algorithms/software (potentially as executables to protect intellectual property) with PNNL, and PNNL will share feedback about the software’s performance on subsets of the NPDS. Such a blind methodology ensures that sensitive information about critical infrastructure remains protected, but the value of the NPDS can be extended to research beyond PNNL. Through iterative feedback, the algorithms may be tweaked to address real-world artifacts. As the NPDS data is temporally and geographically diverse, it may capture features absent in smaller datasets used during the development of the algorithm under test. This report presents lessons learned from applying the blind validation methodology to LineID™, a synchrophasor-based dynamic line rating software developed by Topolonet Corporation. Improvements made to the software through iterative feedback, limitations of the validation methodology, as well as how the limitations of the NPDS affected the evaluation process are discussed. Observations indicate that the proposed validation methodology can be valuable for evaluating other tools in the future.

97 MATHEMATICS AND COMPUTING↗

Electric Vehicle Charging Analytics and Reporting Tool (EV-ChART): Data Format and Preparation Guidance (V.5.0)

The Joint Office of Energy and Transportation maintains the Electric Vehicle Charging Analytics and Reporting Tool (EV-ChART), which provides a centralized hub for submitting electric vehicle (EV) charging infrastructure data directed by the Federal Highway Administration (23 CFR 680.112(a)-(c)). EV-ChART provides a streamlined data submission process and an integrated set of analytic tools, connects to other data sources, and empowers data sharing and access across stakeholders, including the public. Any data shared publicly will be aggregated and anonymized to stay in accordance with 23 CFR 680. This EV-ChART Data Format and Preparation Guidance provides a comprehensive overview of the data reporting requirements as authorized under 23 CFR 680.112(a)-(c)). The guidance is intended to be used alongside the EV-ChART Data Input Template, which defines the tabular data structure that these data submissions must follow. Per 23 CFR 680.112(a)-(c), the annual and quarterly data submissions are required of all National Electric Vehicle Infrastructure (NEVI) Formula Program projects, as well as projects for the construction of publicly accessible EV chargers that are funded with funds made available under Title 23, United States Code, including any EV charging infrastructure project funded with federal funds that is treated as a project on a federal-aid highway. One-time data submissions are required of both the NEVI Formula Program projects and grants awarded under 23 U.S.C. 151(f) for projects that are for EV charging stations located along and designed to serve the users of designated Alternative Fuel Corridors (AFCs). Other information and data required in 23 CFR 680, such as 23 CFR 680.112(d), 23 CFR 680.116(c), and 23 CFR 680.106(a), are not discussed in this guidance.

33 ADVANCED PROPULSION SYSTEMS↗

Self-Lubricating Bushing Testing for John Day Dam Ka plan Turbine Replacement

The goal of this testing program is to find suitable self-lubricated alternatives to replace the in use oil lubricated grooved-bronze bushing in Kaplan runner hubs that are currently being utilized at John Day Dam operated by the U.S. Army Corps of Engineers (USACE). Scaled bushing testing utilizing a Pacific Northwest National Laboratory designed, fabricated, and operated Bushing Test Stand under the direction and funding from USACE determined that self-lubricated bushings tested appear capable of outperforming the current oil-lubricated bronze bushings. Performance tests are focused on durability which includes cyclic load bearing properties and articulations under various loading scenarios. This testing was performed at nominally 1/5 th scale. A scaling sensitivity analysis determined that scaled testing results apply to prototypic size. This report shows the direct comparison between the current baseline bronze bushing to self-lubricated bushings from various manufacturers. Bushing data like coefficient of friction and bushing material loss (associated with bushing wear) was determined. Self-lubricated bushings tested include products manufactured specifically for hydropower applications by the following companies: Trelleborg, Kamatics, Anonymous, and Tenneco/Deva.

42 ENGINEERING↗

Initial Mobility Analysis for ORNL VA-EDH Synthetic Populations

Travel burdens are a major barrier to healthcare access among US Veteran patient populations, particularly those residing in rural areas. Spatial accessibility to points of care for US Veteran populations is commonly assessed in two ways. The first approach uses open data from the US Census to represent collective travel burdens, for example the distance between population-weighted census tract centroids and VHA points of care. The second approach uses restricted-access VHA patient data to measure travel costs (e.g., distance, time) for accessing points of care with respect to geolocated patient addresses and real or approximated transportation networks. While the advantage of the open data approach lies in its reproducibility, it has notable limitations in its tendency to infer individual travel behavior from aggregate population characteristics, a problem known as ecological fallacy. Conversely, while the patient data approach is able to account for individual travel behavior, its ability to account for localized access disparities (e.g., a neighborhood with exceptionally high transportation costs) and patient demographics is limited as protecting individual patient data requires their storage in closed systems with limited capacity for adequately modeling real-world travel patterns or for supplementing patient attributes. Additionally, the patient data approach cannot account for veterans who are not enrolled in the VHA system but who may be eligible for care. These challenges limit the ability to perform “what if” analyses on the effects of place-specific interventions on veteran populations with high access barriers to healthcare. To address these challenges, we explore the application of realistic synthetic populations to examine travel burdens and spatial accessibility issues among veteran patient populations. Synthetic populations provide a virtual, individually-resolved and cross-sectional representation of the veteran patient population that enables investigation of spatial access to points of care in ways in which aggregate data and patient data do not. First, synthetic populations allow one to directly assess how individuals access points of care, from synthesized residential locations to outpatient facilities on real-world transportation networks. Modeling access to points of care at the individual scale addresses the ecological fallacy problem associated with using aggregated census data to represent veteran populations and patterns of movement. Second, synthetic populations provide a means of completely representing an area’s veteran population using only publicly available, anonymized census microdata from the American Community Survey (ACS) to ensure the privacy of real-world individuals. Generating synthetic populations from the ACS also expands descriptive characteristics beyond what patient data typically offers to include socio-demographic, economic, housing, and mobility attributes. More detailed profiles of both VHA patient populations and veterans not enrolled in the VA system will provide a comprehensive picture of groups that may benefit from interventions or outreach. As an initial exercise for using synthetic populations to measure veteran travel burdens to VA care, we apply Oak Ridge National Laboratory’s (ORNL) UrbanPop capability to generate a series of synthetic VHA patient populations for 9 Veterans Integrated Services Networks (VISN) market areas in 9 Census Divisions across the continental United States, which are listed in Table 1. We use UrbanPop to produce synthetic populations for the VISN markets selected for each US Census Division, then assign VA outpatient clinic destinations to synthetic VHA patients based on travel about each VISN market’s road network. To demonstrate using the synthetic populations to evaluate healthcare travel burdens, we compare the time-based impedance between simulated home locations and VA outpatient clinics in each VISN market. We then perform validation exercises on the synthetic populations with respect to neighborhood (block group) demographic composition as well as patient mobility, comparing aggregate origin-destination statistics for the synthetic population to outpatient visits available in restricted patient data from the VA’s Corporate Data Warehouse (CDW) database.

97 MATHEMATICS AND COMPUTING↗

Assessment of BQ-9000 Biodiesel Properties for 2024

This is the eighth in a series of reports documenting the quality of biodiesel from U.S.- and Canadian-based producers that participate in the BQ-9000 program, the biodiesel industry's voluntary quality assurance program. Participants provided monthly data on critical quality parameters for calendar year 2024 with quality data provided to a team of experts, who removed any identifying company information and provided anonymized data to the National Renewable Energy Laboratory (NREL) for statistical analysis. Similar to 2023, data on kinematic viscosity, sulfated ash, distillation temperature, carbon residue, and cetane were collected, as well as individual levels of sodium, potassium, calcium, and magnesium.

09 BIOMASS FUELS↗

Facilitating Data Collection of Maintenance Events to Populate the Hydrogen Component Reliability Database (HyCReD)

The Hydrogen Component Reliability Database (HyCReD) is a collaborative project between the National Renewable Energy Laboratory, the University of Maryland, and hydrogen stakeholders to improve safety and reliability for hydrogen facilities by implementing component reliability data taxonomies that support hydrogen infrastructure failure rate analysis. The project aims to quantify failure rates of hydrogen components through high-quality data collection and analysis on root causes and maintenance needed. HyCReD provides a common database for cataloging hydrogen component failures which exists for reliability research in many other mature industries [2]. The database fills a gap for the hydrogen community by providing a scientifically rigorous approach to quantitative risk assessment (QRA), prognostic health management (PHM), and reliability-centered maintenance (RCM) analysis. High level results will be aggregated and anonymized to protect company sensitive information; detailed results will be used to help address issues of hydrogen components. These advanced analytics will support accelerated deployment of hydrogen infrastructure by enabling better: design and safety of projects (safety codes and standards development), infrastructure reliability and cost (component failure rates, maintenance protocols), and component R&D needs (robust supply chain). A key to a successful HyCReD implementation is facilitating the ease of reporting and data quality in the database that can be used for analysis. Maintenance data was a previously identified gap in initial efforts to populate and validate the database taxonomies [3]. Collection of maintenance data will be instrumental in identifying failure modes and rates, identifying incipient component failures or reduced performance, cataloging best practices for maintenance routines and methods for prognostic health management, and quantifying the risk and effect of different failure modes. Several key priorities are identified for streamlined data collection to achieve quality and detailed failure data: Applicability, Ease of Use, Accessibility, and Information Security. The HyCReD team has now begun deployment of the database to several companies and groups that have signed non-disclosure agreements to facilitate the data collection of failures in industry hydrogen refueling station infrastructure. This paper will provide an update into the process of HyCReD deployment including the development of a coding guide for facility personnel to reference and ensure data quality and consistency from one station to another as well as implementation of contextually dependent data fields of system taxonomy and formatted entries to provide ease of use. The goal is to communicate the lessons learned from the roll-out to technicians and engineers in the field, and the addition of need for high level of security to protect all stakeholders.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Ecobee Donate Your Data 1,000 homes in 2017

This dataset is a subset of the Ecobee Donate Your Data (DYD) dataset. The Ecobee DYD data comprises user-reported metadata (of home and occupant characteristics), and data collected by Ecobee thermostats (reported in 5-minute intervals). Participant data are pulled from the Ecobee servers, and then anonymized to remove any personally identifiable information. This subset selects 1,000 single family homes in four states - California, Texas, New York, and Illinois, and span the entire year of 2017. In addition to the measurements, a metadata JSON file is included to illustrate the high-level contextual information of the dataset. The dataset can be analyzed to understand how a single-family heating, ventilation, and air-conditioning (HVAC) system operates, occupant behavior, and building thermal dynamics.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Blockchain Research and Development Activities Sponsored by the U.S. Department of Energy and Utility Sector

This article provides an in-depth analysis of blockchain research in the energy sector, focusing on projects funded by the U.S. Department of Energy (DOE) and comparing them with industry-funded initiatives. A total of 110 funded activities within the U.S. power industry were successfully tracked and mapped into a newly developed categorization framework. This framework is designed to help research agencies to systematically understand their funded portfolio. Such characterization is expected to help them make effective investments, identify research gaps, measure impact, and advance technological progress to meet national goals. In line with this need, the proposed framework proposes a 2-D categorization matrix to systematically classify blockchain efforts within the energy sector.Under the proposed framework, the Energy System Domain serves as the primary classification dimension, categorizing use cases into 30 distinct applications. The second dimension, Blockchain Properties, captures the specific needs and functionalities provided by Blockchain technology. The aim was to capture blockchain’s applicability and functionality: where and why blockchain? Principles behind the selection of the viewpoint dimensions were carefully defined based on consensus obtained through the Blockchain for Optimized Security and Energy Management (BLOSEM) project. The mapped results show that activities within the Grid Automation, Coordination, and Control (31.8%), Marketplaces and Trading (25.5%), Foundational Blockchain Research (19.1%), and Supply Chain Management (17.3%) domains have been actively pursued to date. The three leading specific use case applications were identified as Transactive Energy Management for Marketplaces and Trading, Asset Management for Supply Chain Management, and Fundamental Blockchain for Foundational Blockchain Research. The Marketplaces and Trading and Retail Services Enablement domains stood out as being favored by industry by a factor greater than 2 (2.3 and 2.6, respectively), yet there seemed to be little to zero investment from DOE. Approximately 76% of the total projects prioritized Immutability, Identity Management, and Decentralization and/or Disintermediation compared to Asset Digitization and/or Tokenization, Automation, and Privacy and/or Anonymity. The greatest discrepancies between DOE and industry were in Asset Digitization and/or Tokenization and Automation. The industry efforts (36% in Asset Digitization/Tokenization and 22% in Automation) was 14 times and 2.4 times, respectively, more intensive than the DOE-sponsored efforts, indicating a significant discrepancy in industry versus government priorities. Overall, quantifying DOE-sponsored projects and industry activities through mapping provides clarity on portfolio investments and opportunities for future research.

24 POWER TRANSMISSION AND DISTRIBUTION↗