Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “geospatial data visualization”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Tool-Based Case Studies on Strategic Deployment of Untapped Micro-Pumped Hydro Storage in Michigan

With most classical hydropower sites already utilized and the global push for rapid integration of renewable energy sources accelerating, there is a critical need to identify alternative energy storage solutions. Pumped hydro energy storage, which accounts for the vast majority of global grid-scale storage, remains one of the most cost-effective and long-duration storage technologies available. Hence, this study presents a novel tool designed to assess the untapped potential of inland lakes and reservoirs for micro-PSH, using Michigan’s relatively flat landscape as a case study due to its extensive but underutilized water infrastructure. To ensure accuracy and reliability, the tool incorporates extensive data gathered from authorized sources, covering more than 420 water facilities and potential reservoirs in the state. The tool evaluates key parameters such as horizontal and vertical distances, volume, and the total storage capacity of each reservoir. Its robust assessment framework integrates these metrics to evaluate each site’s potential. The tool’s intuitive interface and geospatial visualizations support actionable insights for planners and scalable deployment of distributed storage infrastructure.

13 HYDRO ENERGY↗

Hyperspectral remote sensing-based plant community map for region around NGEE-Arctic intensive research watersheds at Seward Peninsula, Alaska, 2017-2019

Using airborne hyperspectral remote sensing data from NASA Airborne Visible-Infrared Imaging Spectrometer- Next Generation (AVIRIS-NG) platforms in a region near NGEE-Arctic intensive watersheds at Seward peninsula of Alaska, high resolution (5m) maps of plant community distribution were developed and included in this data collected. AVIRIS-NG data collected over 2017-2019 period were used to develop deep neural networks, trained using vegetation plot observations collected at NGEE-Arctic watersheds at Kougarok, Council and Teller. A hierarchical vegetation classification scheme consisting of six classes at Level I, and 16 classes at Level II contained in two .txt files were used to developed the plant community maps for the region. Two geospatial raster data files (.tif) at both thematic levels are shared in this data collection. Data files in this collection use Alaska Albers Equal Area projection. Readme files available in three formats (*.html, *.md, *.pdf) and one *.png visualization map.The Next-Generation Ecosystem Experiments: Arctic (NGEE Arctic), was a research effort to reduce uncertainty in Earth System Models by developing a predictive understanding of carbon-rich Arctic ecosystems and feedbacks to climate. NGEE Arctic was supported by the Department of Energy's Office of Biological and Environmental Research.The NGEE Arctic project had two field research sites: 1) located within the Arctic polygonal tundra coastal region on the Barrow Environmental Observatory (BEO) and the North Slope near Utqiagvik (Barrow), Alaska and 2) multiple areas on the discontinuous permafrost region of the Seward Peninsula north of Nome, Alaska.Through observations, experiments, and synthesis with existing datasets, NGEE Arctic provided an enhanced knowledge base for multi-scale modeling and contributed to improved process representation at global pan-Arctic scales within the Department of Energy's Earth system Model (the Energy Exascale Earth System Model, or E3SM), and specifically within the E3SM Land Model component (ELM).

54 ENVIRONMENTAL SCIENCES↗

Legacy surface change analyses of the 1993 Rock Valley earthquake sequence for direct comparison to planned NA-22 underground conventional high-explosive experiments (Source Physics Experiment 3 (SPE3) - RV/DC Task 1.3 FY22 Final Report)

Recent work under two previous phases of the NNSA NA-22 Source Physics Experiment (SPE) have shown that underground chemical high-explosive experiments can produce detectable surface changes that differ in spatial extent and vertical magnitude depending on the geologic media at the site (Schultz-Fellenz et al., 2018; 2020; Crawford et al., 2021). Neither of these two prior phases of SPE identified natural earthquake-related surface effects in the same region that occurred at a similar depth as the explosive experiments for direct comparison. While earthquakes can also produce surface changes that are detectable using remote sensing data analyses, it is expected that the pattern and spatial extent of surface changes would vary between earthquakes and explosions. However, no direct-observed surface-change signature comparison between earthquakes and explosions has ever been performed. The SPE Phase 3 Rock Valley/Direct Comparison (RV/DC) program presents a unique opportunity to investigate and characterize co-occurrence of both earthquakes and explosions. In this project, we worked to address five tasks in a workflow, as follows: 1. Identify and obtain existing high-resolution legacy satellite and aerial imagery as close in time before and after the 1993 Rock Valley earthquake sequence to temporally constrain the analyses. 2. Transform these pre-earthquake and post-earthquake datasets into digital elevation models (DEMs) using geospatial analysis software packages (e.g., ArcGIS, Agisoft Metashape, and Google Earth Engine). 3. Perform DEM differencing analyses to assess and quantify earthquake-related changes from the 1993 sequence, and develop map products that visualize these analyses. 4. Use the analyses from (3) to: (a) assess spatial distribution and magnitude of surface changes due to the 1993 Rock Valley earthquake sequence, and (b) determine parameters of forthcoming, planned explosion-related surface change data collection from sensors mounted on unmanned aerial vehicles (UAVs) (e.g., spatial extent of collection, design and density of survey control, sensors to deploy, forward speed and line spacing of UAV flight lines, flight altitude). 5. Develop a summary report on the analyses, including how the analyses define parameters and identify focus areas for any future surface change analytical field campaigns related to the explosive experiment. Analyzing these legacy data and identifying whether they can detect any surface changes related to the earthquake sequence facilitates opportunities for direct signature comparison of surface change from explosions at one location, which has never previously been performed. Comparing the surface change signatures from a co-located and depth-equivalent earthquake and an explosion could help to advance remote sensing event discrimination techniques. This report summarizes the work completed toward this ambitious goal.

42 ENGINEERING↗

Transportation and Systems Analysis Collaborations in Support of a Federal Consolidated Interim Storage Facility

The U.S. Department of Energy’s Integrated Waste Management (IWM) program under the Office of Nuclear Energy is planning for the future transportation, storage, and eventual disposal of spent nuclear fuel (SNF) and high-level radioactive waste (HLW) from nuclear power plant sites across the United States. To better enable informed decision-making regarding the back end of the nuclear fuel cycle, the IWM program has been sponsoring the development and application of system analysis tools capable of analyzing various options for managing SNF and HLW. With these tools, integrated waste management system (IWMS) architecture analyses are being conducted to support the future deployment of a comprehensive nuclear waste management system that considers all major back-end aspects of the nuclear fuel cycle (i.e., transportation, storage, and disposal). System analyses and assessments typically use these modeling and simulation tools to investigate implications of changes in various assumptions and parameters such as acceptance rates, receipt logic, facility capacities and capabilities, use of standardized canisters, start and stop dates of facilities, etc. The Next Generation System Analysis Model (NGSAM) is an agent-based simulation toolkit that is used for a system-level simulation and analysis focused on SNF management in the United States. An analyst using NGSAM has the ability to define several factors like the number of storage facilities, capacity at each facility, transportation schedules, shipment rates, and other conditions. NGSAM’s primary purpose is to provide a system analyst with capabilities to model the IWMS and gain insights into SNF and HLW management alternatives including the impact of system choices, associated cost estimates, and development of integrated yet flexible approaches. IWM is also developing the Stakeholder Tool for Assessing Radioactive Transportation (START). START is a web-based geospatial tool developed to provide visualization and initial evaluation of transportation options associated with future SNF and HLW shipment planning and operations. This includes characterizing safety, economic, and environmental conditions on and in proximity to shipment origins as well as along prospective transportation routes. Information from the START tool can be presented/shared as maps, graphics, geospatial files, and tabular form to enable easy data representation and export functionality. The system analysis, NGSAM, and START teams had been working closely for several years before formalizing this collaboration. START provides the SNF routing information for use in NGSAM. System analysts use the NGSAM tool to generate results (schedule, costs, infrastructure acquisitions, shipment rates, etc.). Subject-matter experts and system analysts work together to inform how NGSAM should model the waste management system. Specifically, this paper discusses the collaborations between the system analysis, NGSAM, and START teams and their accomplishments over the past year. Some examples of these efforts include calibrating and adding data to the START output files to meet NGSAM needs, gaining a better understanding of START data used in NGSAM, as well as how updates in START data could result in an improved NGSAM analysis. Detailed examples of various tasks that have been performed by the team will be discussed in the full paper. Continued efforts in this direction are expected in the coming years.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Transportation and System Analysis Collaborations in Support of a Federal Consolidated Interim Storage Facility

The U.S. Department of Energy’s Integrated Waste Management (IWM) program under the Office of Nuclear Energy is planning for the future transportation, storage, and eventual disposal of spent nuclear fuel (SNF) and high-level radioactive waste (HLW) from nuclear power plant and waste custodian sites across the United States. To better enable informed decision-making regarding the back end of the nuclear fuel cycle, the IWM program has been sponsoring the development and application of system analysis tools capable of analyzing various options for managing SNF and HLW. With these tools, integrated waste management system (IWMS) architecture analyses are being conducted to support the future deployment of a comprehensive nuclear waste management system that considers all major back-end aspects of the nuclear fuel cycle (i.e., transportation, storage, and disposal). System analyses and assessments typically use these modeling and simulation tools to investigate implications of changes in various assumptions and parameters such as acceptance rates, receipt logic, facility capacities and capabilities, use of standardized canisters, start and stop dates of facilities, etc. The Next Generation System Analysis Model (NGSAM) is an agent-based simulation toolkit that is used for a system-level simulation and analysis focused on SNF management in the United States. NGSAM’s primary purpose is to provide a system analyst with capabilities to model the IWMS and gain insights into SNF and HLW management alternatives including the impact of system choices, associated cost estimates, and development of integrated yet flexible approaches. An analyst using NGSAM can define several factors like the number of storage facilities, capacity at each facility, transportation schedules, shipment rates, and other conditions. IWM is also developing the Stakeholder Tool for Assessing Radioactive Transportation (START). START is a web-based geospatial tool developed to provide the visualization and initial evaluation of transportation options associated with future SNF and HLW shipment planning and operations. This includes characterizing safety, economic, and environmental conditions at and in proximity to shipment origins as well as along prospective transportation routes. Information from the START tool can be presented/shared as maps, graphics, geospatial files, and tabular form to enable easy data representation and export functionality. The system analysis, NGSAM, and START teams have been working closely for several years before formalizing this collaboration. START provides the SNF transportation routing information for use in NGSAM. System analysts use the NGSAM tool to generate results (schedule, costs, infrastructure acquisitions, shipment rates, etc.). Subject-matter experts and system analysts work together to inform how NGSAM should model the waste management system. Specifically, this paper discusses the collaborations between the system analysis, NGSAM, and START teams and their accomplishments over the past year. Some examples of these efforts include calibrating and adding data to the START output files to meet NGSAM needs, gaining a better understanding of START data used in NGSAM as well as how updates in START data could result in an improved NGSAM analysis. Detailed examples of various tasks that have been performed by the team will be discussed in the full paper. Continued efforts in this direction are expected in the coming years.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

The Integration and Mapping of an Open-Source National Well Resource to Inform Geologic Carbon Storage Site Selection and Risk Prevention: The CO2-Locate Database

Geologic carbon storage (GCS) offers a way to capture and permanently store CO₂ from fossil fuel operations in underground geologic structures, aiding in the transition to a carbon-neutral energy economy. However, CO₂ injection sites can experience gas leakage through existing wells that penetrate storage reservoirs, making knowledge of well locations and characteristics crucial for permitting, infrastructure reusability, and risk assessment in GCS. Currently, public wellbore data from state, federal, and tribal entities are inconsistent and fragmented, with gaps and redundancies. To address this, the National Energy Technology Laboratory (NETL) developed CO2-Locate, an open-source, geospatial database and online application. CO2-Locate integrates over 50 data sources from federal, state, and tribal entities, creating a standardized national well database. Funded by the Bipartisan Infrastructure Law, the database is publicly available through the Energy Data eXchange (EDX) and viewable via the CO2-Locate web mapping application. This tool allows users to query, filter, and visualize well data to support GCS planning, permitting, and risk assessments. This presentation covers the methods used to create CO2-Locate, including data acquisition, processing, attribute mapping, and integration, much of which is automated for future updates. The web mapping application and its role in GCS site selection will also be discussed.

Tetteh, Daniel A.↗

Geospatial Characterization of Low-Temperature Heating and Cooling Demand in the United States: Preprint

Geothermal resources at temperatures below 150 degrees C have great potential as energy sources for various direct-use applications including heating and cooling in residential and commercial buildings. This study geospatially quantifies U.S. heating and cooling demand in residential, commercial, and manufacturing sectors; heating demand in the agricultural sector; and cooling demand in data centers at the county level through end-use energy consumption, expenditure, and commissioned power analyses. Heating and cooling demand in the residential sector was estimated using energy consumption data obtained from the U.S. Energy Information Administration accounting for different U.S. climate zones. For commercial sector analysis, the end-use major fuel energy intensity at the census division level was disaggregated to the county level with respect to principal building activities. Heating and cooling demand analysis for the manufacturing sector was based on end-use energy consumption for direct-use total process categorized by the North America Industry Classification System. Fuel expenditures in the U.S. Department of Agriculture Farm Production Expenditures were examined for heating demand analysis in the agricultural sector, particularly for the greenhouse, nursery, and floriculture production category. Lastly, commissioned power for data centers in the United States were explored for cooling demand analysis. Results indicated a significant fraction of U.S. primary energy consumption is used for low-temperature heating and cooling applications. Heating and cooling demand in residential and commercial sectors is significantly affected by the number of housing units and climate zone designations, while heating and cooling demand in manufacturing and agricultural sectors and data centers are mainly dependent on the number of facilities and their locations. Maps were generated visualizing where heating and cooling demand is high and, overlain with geothermal resource maps, can indicate locations where geothermal energy can supply this heating and cooling demand.

cooling demand↗

Geospatial Characterization of Low-Temperature Heating and Cooling Demand in the United States

Geothermal resources at temperatures below 150 degrees C have great potential as energy sources for various direct-use applications including heating and cooling in residential and commercial buildings. This study geospatially quantifies U.S. heating and cooling demand in residential, commercial, and manufacturing sectors; heating demand in the agricultural sector; and cooling demand in data centers at the county level through end-use energy consumption, expenditure, and commissioned power analyses. Heating and cooling demand in the residential sector was estimated using energy consumption data obtained from the U.S. Energy Information Administration accounting for different U.S. climate zones. For commercial sector analysis, the end-use major fuel energy intensity at the census division level was disaggregated to the county level with respect to principal building activities. Heating and cooling demand analysis for the manufacturing sector was based on end-use energy consumption for direct-use total process categorized by the North America Industry Classification System. Fuel expenditures in the U.S. Department of Agriculture Farm Production Expenditures were examined for heating demand analysis in the agricultural sector, particularly for the greenhouse, nursery, and floriculture production category. Lastly, commissioned power for data centers in the United States were explored for cooling demand analysis. Results indicated a significant fraction of U.S. primary energy consumption is used for low-temperature heating and cooling applications. Heating and cooling demand in residential and commercial sectors is significantly affected by the number of housing units and climate zone designations, while heating and cooling demand in manufacturing and agricultural sectors and data centers are mainly dependent on the number of facilities and their locations. Maps were generated visualizing where heating and cooling demand is high and, overlain with geothermal resource maps, can indicate locations where geothermal energy can supply this heating and cooling demand.

cooling demand↗

Interactive web mapping tools and custom subsurface cross-sections for interdisciplinary geologic investigation

Using Python-based geospatial analytics, open-source web mapping technologies, geophysical data models, and subsurface stratigraphy models from the Regional Geology Geologic Framework Model database assembled by Los Alamos National Laboratory, we developed a suite of web-based geologic investigation tools to identify and understand subsurface structures and geophysical properties concerning salt and shale formations within the contiguous United States. Coupled with a web map interface, these tools allow for the interactive visualization of various geologic data and demonstrate the ability to quickly generate custom subsurface cross-sections, borehole charts, and diagrams for azimuthal orientation data. These capabilities were developed for stakeholder and researcher use to facilitate informed decision making for spent nuclear waste disposition. However, these capabilities provide a flexible model for a variety of subsurface investigation needs, and we have demonstrated this flexibility by adapting these tools to meet visualization needs for various subsurface models within a web-based platform.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

Radiation Detection System Integration Using Multiple-Input Multiple-Output Radios

This study describes the application development of multiple-input multiple-output radios to provide persistent mobile ad hoc network (MANET) for the Department of Homeland Security. By using Man Portable Unit (MPU5) fifth generation radios (manufactured by Persistent Systems) with the Android Team Awareness Kit (ATAK), an Android smartphone geospatial infrastructure and military situational awareness application, the Remote Sensing Laboratory has developed a MANET connectivity to monitor deployed nuclear/radiological search operation assets.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Data Fusion for the Development of a Multimodal Freight Transload Facilities Dataset in the U.S.

To withstand the growing demand of commodity volume and its strain on the transportation infrastructure, it is necessary to identify the flow of commodities by route and mode. However, a national multimodal freight routing model does not exist for the U.S. The development of such model requires multiple building blocks, such as virtual representations of roadway, railway, and waterway networks, transload facilities (TFs), and access/egress links. Most of these blocks have a robust database in the U.S., except for the TFs. Here, this paper presents the fusion of dispersed and heterogeneous representations of multimodal TFs into a single, comprehensive, geospatial freight TF dataset. The TF dataset is derived from several sources, including the U.S. Army Corps of Engineers Master Docks Plus, the National Transportation Atlas Database, the Intermodal Association of North America, industry publications, and other public information. First, individual datasets were queried and reconciled. A geocoding/reverse geocoding process was applied to get the best street address and latitude/longitude location for each terminal. Then, duplicate terminals were identified by a fuzzy match algorithm based on terminal name and location, and removed. Validation was performed by visual inspection of random facilities. The main contributions of this work are: a publicly available version of the TF dataset, including facility location and multimodal transfer capability of 9,003 facilities, and an enterprise-version with the same facilities but including commodity handling capabilities. The main purpose of developing the TF dataset is to inform multimodal routing algorithms. The proposed TF dataset allows for credibly modeling the multimodal transfer of commodities within shipment routes.

Commodity Routing↗

MacroSheds: A synthesis of long-term biogeochemical, hydroclimatic, and geospatial data from small watershed ecosystem studies

The US Federal Government supports hundreds of watershed monitoring efforts from which solute fluxes can be calculated. Although instrumentation and methods vary between studies, the data collected and their motivating questions are remarkably similar. Nevertheless, little effort toward their compilation has previously been made. The MacroSheds project has developed a future-friendly system for harmonizing daily time series of streamflow, precipitation, and solute chemistry from 169+ watersheds, and supplementing each with watershed attributes. Here, we describe the breadth of MacroSheds data, and detail the steps involved in rendering each data product. We provide recommendations for usage and discuss when other datasets might be more suitable. The MacroSheds dataset is an unprecedented resource for watershed science, and for hydrology, as a small-watershed supplement to existing collections of streamflow predictors, like CAMELS and GAGES-II. The MacroSheds platform includes a web dashboard for visualization and an R package for data access and analysis.

59 BASIC BIOLOGICAL SCIENCES↗

Vegetation classification map and covariates associated with NEON AOP survey, East River, CO 2018

This package includes geospatial data layers developed to investigate how environmental gradients—specifically topography and near-surface soil properties—drive the spatial arrangement of dominant plant communities in mountainous watersheds. The geospatial products, which support the analysis of these ecological relationships, are derived from airborne hyperspectral and LiDAR datasets acquired by the National Ecological Observatory Network (NEON) Airborne Observation Platform (AOP), in conjunction with an extensive ground field campaign conducted in summer 2018. This work is part of the DOE Watershed Function Science Focus Area (SFA) and features geospatial datasets developed based on observations and ground data collected at East River, Colorado, in collaboration with the National Ecological Observatory Network (NEON) Airborne Observation Platform (AOP) survey in June 2018. Classification Map: - Classification Map (PNG, GeoTIFF): Derived from hyperspectral and LiDAR airborne data using a machine learning approach. - Class Code Mapper (CSV): Associates pixel values with corresponding vegetation/non-vegetation classes. - Classification Reference Data (CSV): Reference data used in the machine learning procedure. LiDAR-Derived Products: - Topographical Metrics (GeoTIFFs): Elevation, slope, curvature, TWI, TPI, solar insolation, and canopy height model (CHM), smoothed with a 5x5 pixel window. Vegetation Indices: - GeoTIFFs of NDVI, NDNI, NDWI: Vegetation indices derived from hyperspectral data. Urban Masks: - Urban Mask (GeoTIFF): Applied to the mapping to convert bare soil classes to urban classes. Software Compatibility: GeoTIFFs: Can be visualized with GIS software or libraries that support GeoTIFF images. CSV Files: Can be opened with any software that handles comma-separated values. The FLMD file provides details and links to the source datasets used to derive the products. The manuscript (in the Method session) provides details on how each product was derived. This work was supported by the Watershed Function Science Focus Area at Lawrence Berkeley National Laboratory funded by the US Department of Energy, Office of Science, Biological and Environmental Research under Contract No. DE-AC02-05CH11231. Update on 2026-03-25: Since the original dataset publication date of 02/28/2020, this package has a new classification map derived by an improved methodology. This update also includes additional ground data that improved the representation of some of the communities. See the methods for further details on what has changed between versions.

2018 NEON and 2025 CHESS Campaigns↗

Optimizing Error-Bounded Lossy Compression for Scientific Data With Diverse Constraints

Vast volumes of data are produced by today's scientific simulations and advanced instruments. These data cannot be stored and transferred efficiently because of limited I/O bandwidth, network speed, and storage capacity. Error-bounded lossy compression can be an effective method for addressing these issues: not only can it significantly reduce data size, but it can also control the data distortion based on user-defined error bounds. In practice, many scientific applications have specific requirements or constraints for lossy compression, in order to guarantee that the reconstructed data are valid for post hoc analysis. For example, some datasets contain irrelevant data that should be isolated in particular and users often have intuition regarding value ranges, geospatial regions, and other data subsets that are crucial for subsequent analysis. Existing state-of-the-art error-bounded lossy compressors, however, do not consider these constraints during compression, resulting in inferior compression ratios with respect to user's post hoc analysis, due to the fact that the data itself provides little or no value for post hoc analysis. In this work we address this issue by proposing an optimized framework that can preserve diverse constraints during the error-bounded lossy compression, e.g., cleaning the irrelevant data, efficiently preserving different precision for multiple value intervals, and allowing users to set diverse precision over both regular and irregular regions. We perform our evaluation on a supercomputer with up to 2,100 cores. Experiments with six real-world applications show that our proposed diverse constraints based error-bounded lossy compressor can obtain a higher visual quality or data fidelity on reconstructed data with the same or even higher compression ratios compared with the traditional state-of-the-art compressor SZ. Furthermore, our experiments also demonstrate very good scalability in compression performance compared with the I/O throughput of the parallel file system.

97 MATHEMATICS AND COMPUTING↗

A Data Processing Pipeline To Extract A Knowledge Graph From Heterogeneous Data For Socio-technical Analysis Of Critical Infrastructure Influence

The code is written in Python and consists of the following pipeline that is implemented in Apache Airflow. This pipeline intends to understand the companies that are directly or indirectly involved with a type of critical infrastructure system at some point in that system's lifecycle. The pipeline takes a configuration file that specifies a list of initial companies to consider, a geographic region of interest, and a set of SEC form types as well as other data sources (e.g. CrunchBase) from which to extract entities and relations. There are four main components to this pipeline as currently implemented: Entity Extraction, Network Construction, Analysis, and Visualization. First, Entity Extraction, is implemented as the `topear-extract_organizations` Apache Airflow workflow. Given an initial query that specifies a geographic region of interest and a time interval, the software will extract CI facilities of interest and organizations that have a direct influence relationship to those facilities (e.g. ownership). During the course of the LDRD, we focused on Electric Vehicle charging stations and this information is available via the Department of Energy (DOE) database on fueling stations maintained by NREL. Within the context of the DOE CESER project, we have focused on Battery Energy Storage Systems (BESS). Second, the Network Extraction component will iteratively construct a social network graph given the set of organizations and people extracted in the previous step. Organizations (and eventually People if desired) are then fed as a query to the `topgear-construct_social_network` Apache Airflow workflow which given a set of initial companies and data sets (e.g. SEC EDGAR form types, OpenCorporates, Crunchbase). This Airflow workflow will iteratively query such data sources to discover relationships with new organizations and people. For example, this module can iteratively query SEC EDGAR for metadata that documents the number of each type of form for the given set of companies and their location. This forms metadata represents a catalog of data sources from SEC EDGAR for the extracted social network knowledge graph. The pipeline then downloads these forms from the website and saves them in a build directory for further processing. These documents are then parsed for entities and relations. Again, we note that in additional to SEC data sources, this step can also pull in information on organizations via API services such as CrunchBase and OpenCorporates or bulk data sources. At the end of this step, the resultant social network, the Critical Infrastructure network, and the edges that encode relationships between organizations and CI facilities, form the Adversarial Socio-Technical Network (ASTN) that informs the analysis. Third, the Analysis component processes these generated ASTN. Previously, that has included the ability to compare prevalence of different vendors for a given infrastructure component type across different regions as well as identify common public and private investors across those vendors. This was demonstrated for EV Charging Stations across several different metropolitan areas within an IEEE PES GridEdge publication. More recently, we have looked at ways to identify infrastructure owners and operators of BESS with the most nameplate capacity across different states as well as other indictors of risk resulting from changes in ownership over time. Finally, the Visualization component consists of an HTML/CSS/JS framework by which users can interact geospatial, operational, and organizational relationships across a given portfolio of Critical Infrastructure facilities. The objective is to provide a library of UI/UX modules that can be repurposed for stakeholder-specific dashboards. All of the modules are related via a common event model that enables UI actions in one view to percolate across the other views.

Weaver, Gabriel [Idaho National Laboratory (INL), ↗

Multidimensional perspectives of geo-epidemiology: from interdisciplinary learning and research to cost–benefit oriented decision-making

Research typically promotes two types of outcomes (inventions and discoveries), which induce a virtuous cycle: something suspected or desired (not previously demonstrated) may become known or feasible once a new tool or procedure is invented and, later, the use of this invention may discover new knowledge. Research also promotes the opposite sequence—from new knowledge to new inventions. This bidirectional process is observed in geo-referenced epidemiology—a field that relates to but may also differ from spatial epidemiology. Geo-epidemiology encompasses several theories and technologies that promote inter/transdisciplinary knowledge integration, education, and research in population health. Based on visual examples derived from geo-referenced studies on epidemics and epizootics, this report demonstrates that this field may extract more (geographically related) information than simple spatial analyses, which then supports more effective and/or less costly interventions. Actual (not simulated) bio-geo-temporal interactions (never captured before the emergence of technologies that analyze geo-referenced data, such as geographical information systems) can now address research questions that relate to several fields, such as Network Theory. Thus, a new opportunity arises before us, which exceeds research: it also demands knowledge integration across disciplines as well as novel educational programs which, to be biomedically and socially justified, should demonstrate cost-effectiveness. Grounded on many bio-temporal-georeferenced examples, this report reviews the literature that supports this hypothesis: novel educational programs that focus on geo-referenced epidemic data may help generate cost-effective policies that prevent or control disease dissemination.

59 BASIC BIOLOGICAL SCIENCES↗

The Pan-Arctic Vegetation Cover (PAVC) database v1.1

The Pan-Arctic Vegetation Cover (PAVC) database contains synthesized field-data observations of vegetation cover from 978 Arctic Alaska plots with observations from 2010 to 2021. The cover datasets contain plot data at both the plant functional type (PFT) and species-level resolution, with standardized PFT definitions and species names. We synthesized publicly available point-intercept and visual estimate plots from the Arctic Vegetation Archive of Alaska, the Alaska Vegetation Plots Database, the North Slope Science Catalog, and the National Ecological Observatory Network; as well as previously unpublished data from the Next-Generation Ecosystem Experiments: Arctic (NGEE Arctic).Users will find four synthesized datasets, 4 associated data descriptor (dd) files, and 1 metadata file in the PAVC database:synthesized_species_fcover.csv contains fractional cover (fcover) for unique accepted species names, where names include vegetation identified at the family, genus, species, subspecies, and variety levels, as well as general functional types across all 5 data sources. The synthesized_species_fcover_dd.csv accompanies this dataset with header information.synthesized_pft_fcover.csv contains fcover for the following PFTs: non-vascular plants with lichen and bryophyte subcategories, trees with deciduous and evergreen subcategories, shrubs with deciduous and evergreen subcategories, graminoids (grasses), and forbs (herbaceous flowering plants) measured as total cover. Litter and “other” cover are also included as total cover. Additional “types” include water and bare ground, which were measured as top cover. The synthesized_pft_fcover_dd.csv accompanies this dataset with header information.species_pft_checklist.csv is a lookup table containing the translation from a dataset species name to an accepted species name and to a PFT. This table can be used to clarify our species to PFT adjudications, and to aid users in assigning their own PFTs. Any issues found in this checklist should be reported in the Issues tab of our github.survey_unit_information.csv contains auxiliary information about the plots synthesized in this database. It contains useful information for filtering plots of interest based on temporal, geospatial, and contextual information about the plot surveys.flmd.csv contains metadata information about each file in the database.This research was performed as a part of the NGEE Arctic project. The NGEE Arctic project was a research effort to reduce uncertainty in Earth System Models by developing a predictive understanding of carbon-rich Arctic ecosystems and feedbacks to climate. NGEE Arctic was supported by the Department of Energy's Office of Biological and Environmental Research.The NGEE Arctic project had two field research sites: 1) located within the Arctic polygonal tundra coastal region on the Barrow Environmental Observatory (BEO) and the North Slope near Utqiagvik (Barrow), Alaska and 2) multiple areas on the discontinuous permafrost region of the Seward Peninsula north of Nome, Alaska.Through observations, experiments, and synthesis with existing datasets, NGEE Arctic provided an enhanced knowledge base for multi-scale modeling and contributed to improved process representation at global pan-Arctic scales within the Department of Energy's Earth system Model (the Energy Exascale Earth System Model, or E3SM), and specifically within the E3SM Land Model component (ELM).

54 ENVIRONMENTAL SCIENCES↗

Extending Shared Socioeconomic Pathways to Future Water Supply In-frastructure Scenarios: A Case Study of San Antonio, TX

Datasets supporting findings and visualization behind Okoye and McManamay (2025) Extending Shared Socioeconomic Pathways to Future Water Supply Infrastructure Scenarios: A Case Study of San Antonio, TX. Environmental Research Communications, DOI: 10.57931/2563186 These datasets contains the results of a site selection analysis for municipal water supply planning across multiple Shared Socioeconomic Pathways (SSPs 1–5) and hard scenario classification of water systems in San Antonio, TX. It includes data at the resolution of individual surface water supply sources (COMIDs) and integrates a wide range of hydrologic, socioeconomic, infrastructural, and scenario-based planning variables. Please refer to the README file provided in Files for more details. Descriptions of the datasets are provided below. Dataset(s) Descriptions: Dataset_SSP1_SSP4.xlsx - Contains data used for site selection optimization under SSP1 to SSP4. This dataset was generated based on multi-indicator computations (e.g., WAI, WQI, ERI, WTC, WIC), scenario demand projections, and resource and spatial constraints, excluding new reservoir values. Dataset_SSP5.xlsx - Used for site optimization under SSP5. Unlike Dataset_SSP1_SSP4, this dataset includes new reservoir features with updated calculations of WAI, WTC, and WIC to reflect the added infrastructure and supply potential. hard_classification.xlsx - Provides the scenario classification output for each site. Includes both the initial scenario classification based on Euclidean Distance and adjusted classifications based on 30% change reduction BAU.zip - Zipped folder of .shp files showing spatially optimized water supply sites per SSP under the Business-As-Usual (BAU) water demand strategy LowGW.zip - Zipped folder of .shp files showing optimized site selections under the Low Groundwater strategy

geospatial↗