Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “community data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Microbial community data from throughfall exclusion experiment: Metadata, SI, community composition, LefSe, and FunGuilR data tables from PARCHED Panama tropical forest soils, 2024-2025

Soil contains more carbon (C) than terrestrial vegetation and the atmosphere combined, with some of the largest terrestrial C stocks in tropical rainforests. Soil microbes decompose organic matter, playing a vital role in the storage or loss of soil C. With climate change, drought conditions are predicted to increase in many tropical regions, including both chronic drying and extended drought, potentially influencing these processes. This project explored the effects of chronic and seasonal drying on soil microbial communities across four distinct tropical forests in a long-term drying experiment. We investigated the effects of a chronic drying manipulation on soil microbial community abundance and variation across different forests and seasons. We also compared findings with previously published data from these forests after short-term drying. This project used soils from a long-term drying experiment established in 2018 across four seasonal lowland forests in Panama. Soils were collected from 0 – 10 cm depths during three seasonal periods in control and drying plots in 2024 and 2025 from a total of 32 plots (n = 4 per forest per treatment). The forests varied in baseline rainfall and soil fertility. We calculated alpha and beta diversity indices and compared taxonomic community composition. We found significant biogeographic variation in microbial diversity and taxonomy, with significant differences across the forests and significant effects of the drying treatment. Metadata and sample IDs are within Metadata_16S.csv and Metadata_ITS.csv. Relative abundance tables of every sample at every season are shown in the Excel workbooks 16S Relative Abundance.xlsx and ITS Relative Abundance.xlsx. They are then also shown in CSV files by each taxonomic level. Linear discriminant analysis effect size (LefSe) tables are shown for the full 16S and ITS datasets (n = 96), subsets for every site at every season (n = 8), and then for the forests with each plot merged by season (n = 8). FunGuildR data table of ITS data is uploaded.

Bacteria↗

Hosting downscaled decision-relevant community data products in ESGF2-US

As regionally-relevant high-resolution Earth system data is increasingly relied upon across scientific, policy, and practitioner communities, there is an urgent need for coordinated and federated infrastructure to store, manage, standardize, and distribute decision-relevant community data products. Substantial effort is required to ensure that these products, which are often critical for regional impact assessments and decision-making, are findable, accessible, interoperable, and reusable. The Earth System Grid Federation US project (ESGF2-US) is addressing this challenge by expanding its open-source, distributed platform to support the hosting and dissemination of downscaled Earth system datasets. This expansion includes aligning new downscaled datasets with developing community standards for metadata and file structure, consistent with existing ESGF archives. This includes ensuring CF-compliance, applying CMORization where appropriate, and developing tools to streamline user access. In this paper, we highlight the technical and coordination work required to bring downscaled data into ESGF2-US and aim to inform the broader Earth system data user community about the growing availability and utility of these curated resources.

ESGF↗

Community Data Contribution to M.E.T.A. with ATF-relevant Hydrided Zr cladding (Coated and Uncoated)

Since the aftermath of the Fukushima Daiichi loss-of-coolant accident, accident-tolerant fuel (ATF) claddings have been developed to improve the coping times in such events. However, the mechanical performance of ATF cladding is crucial in ensuring that it does not negatively impact the mechanical integrity during all other stages of the nuclear fuel cycle, and the validity of the existing safe operating margins must be verified. However, due to the cladding’s tube geometry and textured anisotropy, determination of apparent mechanical properties under certain deformation paths is challenging. In the uniaxial hoop direction, for instance, the measured mechanical stresses include frictional forces caused by loading mandrels or varying deformation paths in the sample during traditional ring tensile testing. This experimental difficulty is exacerbated by the specimen size. However, addressing these challenges enables irradiation separate-effects investigations in which the materials can be inserted in reactors like the High Flux Isotope Reactor, and reducing the material consumption of commercially irradiated material allows for further post-irradiation examinations. Despite the advantages of reduced-scale mechanical testing, any drawbacks from new specimen geometries must be evaluated, and uncertainties from specimen preparation, setup, and analysis methodologies must be understood.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Prediction of Distributed River Sediment Respiration Rates Using Community-Generated Data and Machine Learning

River sediment microbial respiration is a key indicator of ecosystem functioning and the biogeochemical fluxes across this critical zone link surface and subsurface waters. As such, there is tremendous interest in measuring and mapping these respiration rates. Respiration observations are expensive and labor intensive; there is limited data available to the community. An open science, collaborative initiative is collecting samples for respiration rate analysis and multi-scale metadata; this evolving data set is being used for making machine learning (ML) predictions at unsampled sites to help inform continued community engagement. However, it is a challenge to find an optimum configuration for ML models to work with this feature-rich (i.e., 100+ possible input variables) data set. Here, we present results from a two-tiered approach to managing the analysis of this complex data set: (a) a stacked ensemble of models that automatically optimizes hyperparameters and manages the training of many models and (b) feature permutation importance to detect the most important features in the models. The major elements of this workflow are modular, portable, open, and cloud-based thus making this implementation a potential template for other applications. The models developed here predict that sediment organic matter chemistry is one of the most important features for predicting sediment respiration rate. Other larger-scale, important features fall into the categories of climatic, ecological, geological, and fluvial settings. Leveraging these larger-scale features to generate data-driven estimates of river sediment respiration rates reveals spatially consistent but heterogeneous patterns across the river network of the Columbia River Basin.

54 ENVIRONMENTAL SCIENCES↗

The Zooplankton International Geospatial (ZIG) dataset: A global repository of spatiotemporal freshwater zooplankton community composition data to support ecological research

Zooplankton play critical roles in aquatic ecosystem function and food webs. Nevertheless, global syntheses of their abundance and community dynamics are challenging due to methodological differences across monitoring programs, taxonomic inconsistencies, and a lack of standardized metadata. To reconcile these challenges, we assembled, curated, validated, and harmonized the Zooplankton International Geospatial (ZIG) dataset, which includes co-located and contemporaneous zooplankton, water chemistry, and limnological data from 307 lakes and reservoirs. ZIG includes waterbodies from each major lake thermal region and range in size from 0.8-2,805,8600 hectares. Temporal coverage for individual waterbodies ranges between 1-60 years of data (median = 4 years) with sampling from once annually to weekly. ZIG is publicly available and can be used to understand freshwater biodiversity change and its drivers at unprecedented scales, and we consider it to be a cornerstone for future investigations of freshwater biology, chemistry, and ecology.

Figary, Stephanie [Cornell University, Ithaca, NY]↗

The Zooplankton International Geospatial dataset: A global repository of spatiotemporal freshwater zooplankton community composition data from lakes and reservoirs to support ecological research

Zooplankton transfer substantial energy in aquatic food webs and are used as indicators of environmental change. Syntheses of zooplankton community dynamics globally require datasets that span a wide range of environmental gradients; however, these datasets are limited due to methodological differences across programs, taxonomic inconsistencies, and a lack of standardized metadata. To reconcile these challenges, we created the Zooplankton International Geospatial (ZIG) dataset, which includes original zooplankton, water physical and chemical variables, and lake morphometric data from 311 inland lakes and reservoirs. ZIG includes waterbodies ranging in size from 0.005 to 82,100 km2 and spanning broad latitudinal (−47.26 to 64.90) and longitudinal ranges (−165.04 to 176.53). Temporal coverage for individual waterbodies ranges between 1 and 60 yr with sampling frequency ranging from annually to weekly. With its extensive coverage and content, we consider ZIG to be a cornerstone for future investigations of global scale lake biodiversity change.

Figary, Stephanie [Cornell University, Ithaca, NY]↗

UrbanScaping: Community Spatial Data Visualization & Analytics

Evaluating the electrification potential of buildings through retrofitting is crucial for reducing carbon emissions and the carbon footprint of built environments. This study leverages the Automatic Building Energy Modeling (AutoBEM) software, integrating the Model America database to create an urban context-based spatial analysis platform for community engagement and development. We selected Camp Hill Borough, PA, as a case study to analyze building-specific energy performance and evaluate the electrification potential of each building by switching to different Heating, Ventilation, and Air Conditioning (HVAC) systems and measurement components. The simulation results generated by the workflow provide retrofitting suggestions to help mitigate the carbon footprint as well as energy saving statistics of buildings. Additionally, the developed web-based interface serves as a community engagement platform, allowing residents to provide feedback and further develop interactive communication protocols. The outcomes of this project offer a baseline for community electrification planning and contribute to the design of low-carbon communities.

Chowdhury, Shovan [ORNL]↗

TropiRoot 1.0: Database of tropical root characteristics across environments

Tropical ecosystems contain the world's largest biodiversity of vascular plants. Yet, our understanding of tropical functional diversity and its contribution to global diversity patterns is constrained by data availability. This discrepancy underscores an urgent need to bridge data gaps by incorporating comprehensive tropical root data into global datasets. Here, we provide a database of tropical root characteristics. This new database, TropiRoot 1.0, will be instrumental in evaluating an array of hypotheses pertaining to root functional ecology and plant biogeography, both within the tropics and relative to other global biomes. The data compilation was conducted by the TropiRoot Initiative, in partnership with the Fine-Root Ecology Database (FRED) and the Global Root Trait (GRooT) database, Colorado State University (CSU) and the Smithsonian Tropical Research Institute (STRI). Literature search and data extraction were conducted between 2020 and 2024. Literature was identified using Web of Science, Scopus, and complemented using the expert knowledge of members of TropiRoot. To provide broad environmental and geographical distributions, literature searches included root characteristics (traits) across global change drivers, natural gradients, and from different continents. We adopted FRED standardized data columns and streamlined the format to enhance accessibility for data extraction across various user groups. This optimized framework resulted in a smaller, yet comprehensive datasheet. To make the database compatible with other global root trait initiatives, column identification was standardized following the codes provided by FRED. These efforts culminated in data extracted from 104 new sources, resulting in more than 8000 rows of data (either species or community data). Most of the data in TropiRoot 1.0 include root characteristics such as root biomass, morphology, root dynamics, mass fraction, architecture, anatomy, physiology, and root chemistry. This initiative represents a 30% increase in the currently available data for tropical roots in FRED. TropiRoot 1.0 contains root characteristics from 25 different countries, where seven are located in Asia, six in South America, five in Central America and the Caribbean, four in Africa, two in North America, and 1 in Oceania. Due to the volume of data, when ancillary data were available, including soil data, these data were either extracted and included in the database or its availability was recorded in an additional column. Multiple contributors checked the entries for outliers during the collation process to ensure data quality. For text-based observations, we examined all cells to ensure that their content relates to their specific categories. For numerical observations, we ordered each numerical value from least to greatest and plotted the values, checking apparent outliers against the data in their respective sources and correcting or removing incorrect or impossible values. Some data (soil and aboveground) have different columns for the same variable presented in different units, including originally published units, but root characteristics data had units converted to match those reported in FRED. By filling a gap from global databases, TropiRoot 1.0 expands our knowledge of otherwise so far underrepresented regions and our ability to assess global trends. This advancement can be used to improve tropical forest representation in vegetation models. The data are freely available and should be cited when used.

FRED↗

Data from TropiRoot 1.0 database: tropical root characteristics across environments

TropiRoot 1.0 is a new tropical root database with root characteristics across environment gradients. It has data extracted from 104 new sources, resulting in more than 8000 rows of data (either species or community data). Most of the data in TropiRoot 1.0 includes root characteristics such as root biomass, morphology, root dynamics, mass fraction, architecture, anatomy, physiology and root chemistry. This initiative represents an approximately 30% increase in the currently available data for tropical roots in the Fine Root Ecology Database (FRED). TropiRoot 1.0, contains root characteristics from 25 different countries where seven are located in Asia, six in South America, five in Central America and the Caribbean, four in Africa, two in North America, and 1 in Oceania. Due to the volume of data, when ancillary data was available, including soil data, these data was either extracted and included in the database or their availability was recorded in an additional column. Multiple contributors checked the entries for outliers during the collation process to ensure data quality. For text-based observations, we examined all cells to ensure that their content relates to their specific categories. For numerical observations, we ordered each numerical value from least to greatest and plotted the values, checking apparent outliers against the data in their respective sources and correcting or removing incorrect or impossible values. Some data (soil and aboveground) have different columns for the same variable presented in different units, including originally published units, but root characteristics data had units converted to match the ones reported in FRED. By filling a gap from global databases, TropiRoot 1.0 expands our knowledge of otherwise so far underrepresented regions, and our ability to assess global trends. This advancement can be used to improve tropical forest representation in vegetation models.

54 ENVIRONMENTAL SCIENCES↗

Data-driven Community-centered Resilient Assessment and Planning Toolkit for Nexus of Energy and Water (DCRAPT-NEW)

Urban areas, including Detroit and Pittsburgh, have suffered significant dual outages of the electrical and water infrastructure in the past decade due, in part, to the increasing number of extreme weather events. With increasing temperatures and rainfall intensity, these regions need to prepare for increasing extreme events through community-based energy and water resilience analysis, planning, and enhancement. This project developed a suite of open-source, open-access, community-centered, data-driven assessment and distributed energy resource (DER) and planning tools for energy and water resilience enhancement in urban areas. Through establishing a multi-level community awareness and engagement mechanism and a comprehensive collection of power outage and flooding data, an innovative group of community energy and water resilience assessment and planning tools have been developed for a wide range of users with differing and variable sets of data available to them. The developed tools include (1) DOE EAGLE-I data-driven, deep-learning assisted resilience assessment and DER planning tools at the county level with socioeconomic factors incorporated; (2) Utility annual power outage data-driven tools for long term resilience assessment and DER planning and 15-min power outage data-driven tools for short term resilience assessment and planning; (3) Detailed engineering tools for energy and water systems resilience assessment and planning when the system topology and component fragility curves are available; (4) Alternative Resiliency Metric Calculation that extracts and separates outage and restoration processes; and (5) Co-optimization tools that evaluate the resilience of the power and sewage system and allow users to conduct joint planning with energy and wastewater systems. The developed tools provide planners, decision-makers, and stakeholders with powerful capabilities to systematically evaluate system/community resilience and optimal and actionable guidance for enhancing resilience while prioritizing DER investments. The tools have been used and validated in Detroit and Pittsburgh and can be used in other areas of the nation. In addition, this project will (1) advance the knowledge and applications of machine-learning methods in analyzing and fusing different layers of information and generating meaningful data points such as generating rare weather events; (2) significantly improve the energy and water resilience of the identified communities in Detroit and Pittsburgh and prepare for more frequent and severe weather conditions; (3) help communities assess extreme weather event impacts and address short-term and long-term resilience-related issues The developed tools have been made public via GitHub and demonstrated to community stakeholders and utility companies via the two annual workshops and numerous community engagement meetings. The project outcomes are also disseminated through publications in various journals and conference proceedings, and presentations at top conferences.

13 HYDRO ENERGY↗

Sharing the Sun Community Solar Project Data

This database represents a list of community solar projects, complete and pending, identified through various sources. The dataset is updated multiple times per year. The current version is the first file located below. Previous versions of the dataset published before June of 2024 can be found in the dataset below labeled “ARCHIVE_Sharing the Sun Community Solar Project Data_Before 06.24.“ The list has been reviewed but errors may exist, and the list may not be comprehensive. Errors in the sources e.g. press releases may be duplicated in the list. Blank spaces represent missing information. NLR invites input to improve the database including, to correct erroneous information, add missing projects, fill in missing information, and remove inactive projects. Updated information can be submitted to Sudha Kannan ( sudha.kannan@nlr.gov ).

14 SOLAR ENERGY↗

Maps of growing season gross primary production and net ecosystem exchange for Council Road Mile Marker 71, Seward Peninsula, Alaska, [2017-2023]

This data archive is in support of the Next-Generation Ecosystem Experiments in the Arctic (NGEE Arctic) publication "Integrating Characteristic Arctic Vegetation in a Land Surface Model Improves Representation of Carbon Dynamics Across a Tundra Landscape", by Murphy et al. (2025a). Murphy et al. (2025a) evaluated whether incorporating observed Arctic vegetation heterogeneity into ELM, the land model of the Department of Energy’s Energy Exascale Earth System Model (E3SM), improved simulations of tundra carbon cycling. The associated model archive can be found at Murphy et al. (2025b). The study focused on the spatial patterns and net landscape-level growing season productivity and carbon uptake. As part of this evaluation, observationally derived maps of average growing season (June–August) net ecosystem exchange (NEE) and gross primary production (GPP) were developed for the same domain. These maps, which form the dataset described here, integrate eddy covariance flux tower, remote sensing, and vegetation community data to provide spatially explicit benchmarks for model evaluation. The maps provide spatially explicit estimates of average growing season NEE and GPP across 13 tundra vegetation communities within the study domain. By combining flux tower observations with Airborne Visible-Infrared Imaging Spectrometer-Next Generation (AVIRIS-NG) hyperspectral imagery and drone-based normalized difference vegetation index (NDVI), these maps capture the heterogeneity of carbon fluxes associated with different Arctic vegetation types. While they represent average seasonal conditions rather than interannual variability, the maps provide a unique dataset for evaluating model performance, comparing vegetation community contributions to landscape-scale carbon cycling, and supporting regional analyses of Arctic carbon dynamics. This data archive contains 5 m resolution maps of vegetation communities, vegetation community average growing season GPP, and vegetation community average growing season NEE (three *.tif files), a User’s Guide (*pdf file), and Table 1 of the User’s Guide displaying vegetation community coverage and average growing season NEE and GPP values (*.csv file).

Murphy, Bailey [ORNL] (ORCID:0000000203995221)↗

A 1-year study on SARS-CoV-2 variant shifts in wastewater using dPCR: comparison with clinical and GISAID data

Wastewater testing can be used to monitor SARS-CoV-2 infections in communities. Data from PCR-based wastewater testing are usually available to public health authorities within 5–7 days after excreta and other body fluids enter the sewer. While PCR-based methods can accurately detect and quantify SARS-CoV-2, sequencing-based methods are usually required to distinguish between variants, delaying the results and adding cost to the process. We developed and assessed a novel, customizable digital PCR (dPCR)-based genotyping method for SARS-CoV-2 variant detection in wastewater, which is more cost-effective, faster, and more accessible than sequencing. This approach was applied to more than 1,400 wastewater samples

Wilton, Rose↗

Effects of fire and fire-induced changes in soil properties on post-burn soil respiration

Boreal forests cover vast areas of land in the northern hemisphere and store large amounts of carbon (C) both aboveground and belowground. Wildfires, which are a primary ecosystem disturbance of boreal forests, affect soil C via combustion and transformation of organic matter during the fire itself and via changes in plant growth and microbial activity post-fire. Wildfire regimes in many areas of the boreal forests of North America are shifting towards more frequent and severe fires driven by changing climate. As wildfire regimes shift and the effects of fire on belowground microbial community composition are becoming clearer, there is a need to link fire-induced changes in soil properties to changes in microbial functions, such as respiration, in order to better predict the impact of future fires on C cycling. We used laboratory burns to simulate boreal crown fires on both organic-rich and sandy soil cores collected from Wood Buffalo National Park, Alberta, Canada, to measure the effects of burning on soil properties including pH, total C, and total nitrogen (N). We used 70-day soil incubations and two-pool exponential decay models to characterize the impacts of burning and its resulting changes in soil properties on soil respiration. Laboratory burns successfully captured a range of soil temperatures that were realistic for natural wildfire events. We found that burning increased pH and caused small decreases in C:N in organic soil. Overall, respiration per gram total (post-burn) C in burned soil cores was 16% lower than in corresponding unburned control cores, indicating that soil C lost during a burn may be partially offset by burn-induced decreases in respiration rates. Simultaneously, burning altered how remaining C cycled, causing an increase in the proportion of C represented in the modeled slow-cycling vs. fast-cycling C pool as well as an increase in fast-cycling C decomposition rates. Together, our findings imply that C storage in boreal forests following wildfires will be driven by the combination of C losses during the fire itself as well as fire-induced changes to the soil C pool that modulate post-fire respiration rates. Moving forward, we will pair these results with soil microbial community data to understand how fire-induced changes in microbial community composition may influence respiration.

54 ENVIRONMENTAL SCIENCES↗

Hydropower potential derived from streamflow extremes for Alaska, USA

Alaska is an expansive region known for its abundant natural resources, including thousands of miles of streams and rivers. These rivers represent potential opportunities for future hydropower development that could provide reliable energy supply for local communities. There is limited long-term high temporal resolution streamflow data available for the region, making data-driven estimates of potential hydropower and its variability across the state challenging. This study provides a novel data-driven approach for hydropower capacity estimation across Alaska. We use supervised machine learning to develop a relationship between the daily and peak flow duration curves in order to augment the size of our dataset from 44 sites to 67 sites. We perform a stochastic hydropower estimation across the 67 sites and identify approximately 1000 MW of total potential hydropower capacity distributed across these sites. Our study provides the first step towards more comprehensive hydropower estimation for this critical region, highlighting the need for future work integrating high-resolution spatial data, community needs, and economic constraints in estimates of potential hydropower development in Alaska.

Hydropower↗

A Parameter-masked Mock Data Challenge for Beyond-two-point Galaxy Clustering Statistics

The past few years have seen the emergence of a wide array of novel techniques for analyzing high-precision data from upcoming galaxy surveys, which aim to extend the statistical analysis of galaxy clustering data beyond the linear regime and the canonical two-point (2pt) statistics. We test and benchmark some of these new techniques in a community data challenge named “Beyond-2pt,” initiated during the Aspen 2022 Summer Program “Large-Scale Structure Cosmology beyond 2-Point Statistics,” whose first round of results we present here. The challenge data set consists of high-precision mock galaxy catalogs for clustering in real space, in redshift space, and on a light cone. Participants in the challenge have developed end-to-end pipelines to analyze mock catalogs and extract unknown (“masked”) cosmological parameters of the underlying ΛCDM models with their methods. The methods represented are density-split clustering, nearest neighbor statistics, BACCO power spectrum emulator, void statistics, LEFTfield field-level inference using effective field theory (EFT), and joint power spectrum and bispectrum analyses using both EFT and simulation-based inference. In this work, we review the results of the challenge, focusing on problems solved, lessons learned, and future research needed to perfect the emerging beyond-2pt approaches. The unbiased parameter recovery demonstrated in this challenge by multiple statistics and the associated modeling and inference frameworks supports the credibility of cosmology constraints from these methods. The challenge data set is publicly available, and we welcome future submissions from methods that are not yet represented.

Krause, Elisabeth [Univ. of Arizona, Tucson, AZ (U↗

Transportability of exogenous microbial community correlates with interwell connectivity in deep aquifers

Subsurface resource engineering operations often utilize continuous injection of externally-sourced water into geological reservoirs for formation pressure maintenance, resource recovery or energy/waste storage. Such injected water generally contains naturally occurring microbes. Little is known, however, about how the injectate microbes transport through geological media as a community, how such transportability is affected by injector-producer connectivity, and whether such knowledge can be utilized for flowpath characterization. In this study, we analyzed daily-to-weekly timeseries microbial community data from the injected- and produced-fluids of a ten-month flow test at a deep, well-characterized engineered aquifer. We found that the injectate microbial community was distinct from the indigenous community at the amplicon sequence variant (ASV) level, and that the transportability of injectate community towards a given producer, quantified by an “nASV-Overlap” metric we propose, had strong and significant positive correlation with known injector-producer connectivities at our site. This suggests that the better the connectivity, the higher the probability for more injectate species to flow through the interwell region and arrive at a producer. Because interwell connectivity is an important yet usually unknown parameter in subsurface resource engineering, such correlation in turn points to nASV-Overlap as a useful indicator of interwell connectivity for aquifer characterization and long-term monitoring. Based on our findings, an nASV-Overlap-based microbial tracing approach was developed for characterizing and monitoring the relative connectivities across multiple producers with a given injector. A side-by-side comparison between the new nASV-Overlap approach and traditional artificial tracer methods is presented, and their respective strengths and limitations are discussed.

Deep biosphere↗