Engineering PapersSearch

SEARCH · Engineering Papers

Results for “compositional data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

The Zooplankton International Geospatial dataset: A global repository of spatiotemporal freshwater zooplankton community composition data from lakes and reservoirs to support ecological research

Zooplankton transfer substantial energy in aquatic food webs and are used as indicators of environmental change. Syntheses of zooplankton community dynamics globally require datasets that span a wide range of environmental gradients; however, these datasets are limited due to methodological differences across programs, taxonomic inconsistencies, and a lack of standardized metadata. To reconcile these challenges, we created the Zooplankton International Geospatial (ZIG) dataset, which includes original zooplankton, water physical and chemical variables, and lake morphometric data from 311 inland lakes and reservoirs. ZIG includes waterbodies ranging in size from 0.005 to 82,100 km2 and spanning broad latitudinal (−47.26 to 64.90) and longitudinal ranges (−165.04 to 176.53). Temporal coverage for individual waterbodies ranges between 1 and 60 yr with sampling frequency ranging from annually to weekly. With its extensive coverage and content, we consider ZIG to be a cornerstone for future investigations of global scale lake biodiversity change.

Figary, Stephanie [Cornell University, Ithaca, NY]

The Zooplankton International Geospatial (ZIG) dataset: A global repository of spatiotemporal freshwater zooplankton community composition data to support ecological research

Zooplankton play critical roles in aquatic ecosystem function and food webs. Nevertheless, global syntheses of their abundance and community dynamics are challenging due to methodological differences across monitoring programs, taxonomic inconsistencies, and a lack of standardized metadata. To reconcile these challenges, we assembled, curated, validated, and harmonized the Zooplankton International Geospatial (ZIG) dataset, which includes co-located and contemporaneous zooplankton, water chemistry, and limnological data from 307 lakes and reservoirs. ZIG includes waterbodies from each major lake thermal region and range in size from 0.8-2,805,8600 hectares. Temporal coverage for individual waterbodies ranges between 1-60 years of data (median = 4 years) with sampling from once annually to weekly. ZIG is publicly available and can be used to understand freshwater biodiversity change and its drivers at unprecedented scales, and we consider it to be a cornerstone for future investigations of freshwater biology, chemistry, and ecology.

Figary, Stephanie [Cornell University, Ithaca, NY]

Data-Driven Compositional Optimization in Misspecified Regimes

With a manifold growth in the scale and intricacy of systems, the challenges of parametric misspecification become pronounced. These concerns are further exacerbated in compositional settings, which emerge in problems complicated by modeling risk and robustness. In “Data-Driven Compositional Optimization in Misspecified Regimes,” the authors consider the resolution of compositional stochastic optimization problems, plagued by parametric misspecification. In considering settings where such misspecification may be resolved via a parallel learning process, the authors develop schemes that can contend with diverse forms of risk, dynamics, and nonconvexity. They provide asymptotic and rate guarantees for unaccelerated and accelerated schemes for convex, strongly convex, and nonconvex problems in a two-level regime with extensions to the multilevel setting. Surprisingly, the nonasymptotic rate guarantees show no degradation from the rate statements obtained in a correctly specified regime and the schemes achieve optimal (or near-optimal) sample complexities for general T-level strongly convex and nonconvex compositional problems.

Business & Economics

NuMaCo

The Nuclear Material and Composition (NuMaCo) server is a web-based platform for storing, managing, and serving material composition data. It includes the SCALE Standard Composition Library and the PNNL Compendium, along with their associated references and uncertainty data. NuMaCo also provides tools to generate, compare, and validate material input definitions for both SCALE and MCNP.

Skutnik, Steve (000000016441135X)

Predictive Assessment of the Chemical Composition of Coal Ash in Reserve at U.S. Disposal Sites

In the United States, more than 2 Gt of coal combustion residuals (i.e., coal ash) are stored in hundreds of disposal units. Recent federal regulations mandate the closure or retrofitting of most coal ash impoundments, presenting significant challenges for waste management. These regulatory pressures also present opportunities to reuse coal ash. However, the quality and quantity of discarded coal ash across the U.S. are not well known, even though this information is crucial for spurring its reuse for conventional and new material applications. This study describes a predictive model for the major element composition of coal ash in reserve at disposal sites of major U.S. coal-fired power plants. This model was constructed from coal purchase records of 705 power stations from 1973 to 2022 and was trained on coal ash composition data, showing that coal ash elemental composition is strongly associated with the source of feedstock coal. The model showed regional shifts in the major element contents of ash produced by power plants in the last 50 years, particularly for calcium and iron (expressed as %CaO and %Fe2O3), as power stations changed their source of coal over this time frame. Our approach enables an estimation of chemical composition for ash stored in waste impoundments at individual power stations. Such information can help to delineate the regional market resource potential of supplementary cements for concrete and other material innovations that would utilize coal ash harvested from disposal sites across the U.S.

01 COAL, LIGNITE, AND PEAT

Characterization of Arsenic and Selenium in Coal Fly Ash to Improve Evaluations for Disposal and Reuse Potential (Final Technical Report)

Coal fly ash is a high volume waste material that is discarded in landfills and surface water impoundments across the U.S. and is also widely recycled for a variety of applications. The leaching of potential of contaminants of concern, such as arsenic (As) and selenium (Se), is often the driver of risk assessments for coal ash disposal and reuse. The extent of leachable As and Se depends on several factors related to environmental conditions and fly ash characteristics. Previous studies employed various methods to delineate the concentration, chemical form, and distribution of As and Se in fly ash materials. However, few studies have attempted to directly correlate these properties to mobilization parameters relevant to disposal and reuse. Instead, the coal residuals industries often rely upon standardized leaching protocols that can be laborious or involve hazardous chemicals. The goals of the project were to: 1) Develop and evaluate a characterization protocol that can be used to screen fly ash samples for leachability of As and Se; 2) Characterize As, Se, and associated constituents of fly ash particles at multiple length scales (nanometer to micrometer) to determine if elemental associations differ as a function of the resolution of characterization; and 3) Establish a predictive model for the chemical composition of coal ash produced annually at major U.S. coal fired power facilities on 50-year national coal supply records. For the first objective, we performed leaching experiments with 52 fly ash samples collected from 15 different U.S. power plants and representing coal feedstocks from the three major domestic coal regions. For this work, we assessed the mobilization potential of As and Se in fly ash based on standardized leaching protocols and performed multivariate and lasso regression analyses to explore correlations of leachable As and Se contents with characteristics such as major element contents, loss on ignition (LOI) and pH. The results of regression models indicated that major elements (Fe, Ca, Al) for a wide range of fly ashes can serve as predictor variables for the leaching potential of As, but not for Se. LOI and pH were not important predictive variables in the models. Both regression approaches resulted in relatively strong fits for leachable As (correlation coefficient R 2 = 0.78 for both models) compared to models for leachable Se (R 2 = 0.49). Overall, these results suggest that correlation models combined with on-site elemental analysis with portable analyzers may enable a screening method for leachable As in coal ash. For the second objective, we utilized nanoscale 2-D imaging (30-50 nm spot size) with the Hard X-ray Nanoprobe (HXN) in combination with microprobe X-ray capabilities (~5 µm resolution) to determine As and Se elemental associations in fly ash particles. Speciation of As and Se was also measured at the nano- to microscale with X-ray absorption spectroscopy. The enhanced resolution of HXN showed As and Se that were diffusely located around or comingled with Ca- and Fe-rich particles. The results also showed nanoparticles of Se attached to the surface of fly ash grains. Overall, a comparison of As and Se species across scales highlights the heterogeneity and complexity of chemical associations for these trace elements of concern in coal fly ash. For the final objective, we developed a predictive model for major element composition of coal ash in reserve at disposal sites of major U.S. coal fired power plants. This model was constructed from coal purchase records of 705 power stations from 1973-2022 and was trained on coal ash composition data showing that coal ash elemental composition is strongly associated with the source of feedstock coal. The model showed regional shifts in the major element contents of ash produced by power plants in the last 50 years, particularly for calcium and iron (expressed as %CaO and %Fe 2 O 3 ), as coal-fired power stations changed their source of coal over this time frame. Our approach enables an estimation of coal ash chemical composition that is stored in waste impoundments at individual power stations. Such information can help delineate the regional market potential for material applications that would utilize coal ash harvested from disposal sites across the U.S.

01 COAL, LIGNITE, AND PEAT

Developing Open-Source Tools for Increasing the Efficiency of Synthetic Aviation Turbine Fuel Certification Process

FuelLib is an open-source Python-based fuel library, developed by NREL, that leverages the group contribution method (GCM) of [1] to systematically estimate the thermodynamic and transport properties of hydrocarbon fuels. FuelLib predicts these properties based on the molecular structure of individual compounds or compound families, using weight percentages of a fuel's composition, typically measured using techniques such as gas chromatography (GC). FuelLib enables property estimation over a wide range of temperatures and pressures of multi-component fuels in the absence of detailed molecular composition data, making it particularly valuable for complex fuel mixtures where detailed experimental characterization of fuel composition is unavailable. These capabilities contribute directly to synthetic aviation turbine fuels (SATF) development, supporting the short-term American Society for Testing and Materials (ASTM) qualification of drop-in fuels while potentially expanding ASTM boundaries to certify a broader range of fuels.

33 ADVANCED PROPULSION SYSTEMS

National Energy Water Treatment & Speciation (NEWTS): A Water & Critical Mineral Database and Dashboard

The scarcity of water resources, the need for beneficial water reuse, and the challenges of wastewater treatment are becoming increasingly pressing in economic, social, and environmental domains. Addressing these concerns requires effective treatment strategies to manage wastewater streams and tackle environmental and economic issues. Furthermore, the recovery of critical minerals from the waste streams associated with energy production holds the promise of offsetting treatment costs and securing local sources of valuable minerals. However, relevant data on these waste streams are dispersed and challenging to locate. The process of ingesting such data into modeling software often involves multiple steps, requiring data restructuring to meet software-input requirements. The non-standardized reporting of water data makes data aggregation and reformatting a time-consuming process. Additionally, essential attributes necessary for modeling water treatment and mineral scale formation are frequently missing. Moreover, data gaps vary depending on the region of interest. Consequently, there is a pressing need for high-quality energy-water composition data that can be easily imported into water chemistry modeling software. To address this need, the National Energy Technology Laboratory has created the National Energy Water Treatment and Speciation (NEWTS) Database and Dashboard—a free online tool catering to community leaders and water researchers. NEWTS facilitates a comprehensive understanding of the composition of energy-related wastewater streams in the United States. The datasets provide detailed concentrations and speciation of major and minor aqueous compounds in energy-related wastewater streams, including power plant leachate, acid mine drainage, brackish water, and oil and gas produced water across the United States. Many of the aqueous species are critical minerals (Li, REEs) in high demand to modernize the world’s energy infrastructure. Many of the datasets also contain volumetric flow-rates needed to model the treatment and reuse scenarios in advanced aqueous chemistry software programs. The NEWTS Database and Dashboard offer public access to hitherto challenging-to-access datasets, presented in a standardized format that is tailored for easy input into aqueous chemistry modeling software. By performing the work needed to transform dispersed, disparate data sources into unified, model-ready datasets, NEWTS serves as an essential resource in advancing water treatment research and sustainable water resource management.

produced water management

Extension of Complex Refractive Index Measurements to the Near-Infrared for Liquids: Methodology and Uncertainty Analysis

Optical identification of liquid droplets, aerosols, or thin films is important for many applications. While reference spectra are sometimes available for such measurements, they are not always applicable to the observed spectrum or the given sample morphology. Reference spectra for many forms can be modeled, however, if the n/k vectors (real and imaginary refractive indices) are available. In previous work we have reported protocols to determine the n/k vectors for dozens of liquids, primarily in the mid-infrared (MIR) spectral range from 7500 to 400 cm –1 . In this work we extend the spectral range into the near-infrared (NIR) region, demonstrating a method to measure and merge the data sets to create composite n/k data ranging from 10 000 to 400 cm –1 (1.0 to 25 µm) with absorbance fidelity spanning over four orders of magnitude, and vastly improved signal-to-noise in the NIR. The precision of the composite data is evaluated for three different liquids, focusing primarily on the steps for converting the raw absorbance spectra to k values. The variability in both MIR and NIR data as well as in the final n/k vectors is also investigated for several liquids. For typical liquids, the overall variability (reported as 2σ) in the final n and k-vectors is determined to be ∼0.4% and 3%, respectively. Finally, the derived n/k data are used to calculate absorbance spectra for aerosol droplets, showing marginal variability due to the typical measurement errors in the final n/k vectors.

47 OTHER INSTRUMENTATION

Revealing EDL-driven reduction mechanisms in binary, ternary, and quaternary fluorinated electrolytes via an integrated MD–DFT–ML framework

Accurately predicting solid electrolyte interphase (SEI) formation requires explicitly resolving the electric double layer (EDL) structure, which deviates significantly from that of the bulk electrolyte. Although an established molecular dynamics (MD) and Density Functional Theory (DFT) framework can model SEI formation by evaluating reduction reactions of local clusters in the EDL, it suffers from a combinatorial computational bottleneck. To overcome this limitation, we introduce a machine-learning-accelerated simulation workflow (MD–DFT–ML), integrating a gradient-boosted regression model trained on EDL composition data to efficiently predict reduction potentials. We apply this framework to seven fluorinated electrolytes comprising fluorinated anions, a fluorinated ester solvent, two types of diluent (ion-solvating ester vs. non-solvating ether), and an FEC additive. The analysis shows that the EDL selectively accumulates cation-binding species; consequently, the non–cation-binding ether diluent rarely enters the EDL and makes minimal contributions to SEI formation. DFT calculations on statistically representative EDL clusters provide reduction potentials and fluorine-release pathways, while the ML model, which substantially reduces the DFT workload, predicts cluster reduction energies with a mean absolute error of 0.1 eV. The combined MD–DFT–ML approach also quantifies contributions from different sources to LiF formation in the SEI. This methodology establishes a generalizable route for multiscale modeling electrolyte and interphase design for next-generation electrochemical energy-storage systems.

DFT-MD-ML workflow

Glass Property-Composition Models Update for use in Direct Feed High-Level Waste Flowsheet Development

A set of preliminary glass property models and constraints were developed and augmented by models from literature for use in design of direct-feed high-level waste (DFHLW) glasses for flowsheet evaluation, testing, and design of the Tank Waste Treatment and Immobilization Plant (WTP) high-level waste (HLW) Facility. These models and constraints are meant to be used as a place-holder while glass property-composition data gaps are filled and final plant operating models are developed. This report describes the motivation and intended use of the models, the compilation of data, model fitting and selection, methods to apply the models and constraints in glass design and offers example calculations demonstrating their intended use.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W

Nepheline constraint for hanford HLW glass

The present work was intended to expand the available HLW glass property-composition data relating to nepheline formation on CCC heat treatment and to develop improved models to predict nepheline formation that would allow processing of high waste loading high-Al HLW glass formulations at the WTP.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W

Glass Property-Composition Models Update for use in Direct Feed High-Level Waste Flowsheet Development: EWG3.0

A set of preliminary glass property models and constraints were developed and augmented by models from literature for use in design of Direct Feed High-Level Waste glasses for flowsheet evaluation, testing, and design of the High-Level Waste Facility at the Hanford Waste Treatment and Immobilization Plant. These models and constraints are meant to be used as a placeholder while glass property-composition data gaps are filled and final plant operating models are developed. This report describes the motivation and intended use of the models, the compilation of data, model fitting and selection, and methods to apply the models and constraints in glass design, and offers example calculations demonstrating their intended use.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W

Glass Property-Composition Models Update for use in Direct Feed High-Level Waste Flowsheet Development: EWG2.6

A set of preliminary glass property models and constraints were developed and augmented by models from literature for use in design of direct-feed high-level waste (DFHLW) glasses for flowsheet evaluation, testing, and design of the Waste Treatment and Immobilization Plant (WTP) high-level waste (HLW) Facility. These models and constraints are meant to be used as a place-holder while glass property-composition data gaps are filled and final plant operating models are developed. This report describes the motivation and intended use of the models, the compilation of data, model fitting and selection, methods to apply the models and constraints in glass design and offers example calculations demonstrating their intended use.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W

Particle composition measurements for ultrafine particles collected at the EPCAPE Mount Soledad site from 04-27-2023 to 06-13-2023 using a Thermal Desorption Chemical Ionization Mass Spectrometer

The dataset contains particle composition data for both the positive and negative reagent ion modes of the Thermal Desorption Chemical Ionization Mass Spectrometer (TDCIMS). The dataset is split into two directories: one for particles with diameters of 30 nm and the other for particles with diameters of less than 100 nm. The positive reagent ion mode uses H3O+ as the reagent ion, and ionization usually occurs through hydrogen addition. The negative mode uses O2- as the reagent ion. Negative mode ionization generally occurs through hydrogen abstraction, but O2- addition is also possible. Ion concentrations were normalized to total ion counts, and unknown ions were then removed from the data. The name of each ion fraction time series includes the mass to charge ratio and the chemical formula for the ion. Time is recorded in seconds since 1/1/1904. The time zone is UTC.

54 ENVIRONMENTAL SCIENCES

Microbial community data from throughfall exclusion experiment: Metadata, SI, community composition, LefSe, and FunGuilR data tables from PARCHED Panama tropical forest soils, 2024-2025

Soil contains more carbon (C) than terrestrial vegetation and the atmosphere combined, with some of the largest terrestrial C stocks in tropical rainforests. Soil microbes decompose organic matter, playing a vital role in the storage or loss of soil C. With climate change, drought conditions are predicted to increase in many tropical regions, including both chronic drying and extended drought, potentially influencing these processes. This project explored the effects of chronic and seasonal drying on soil microbial communities across four distinct tropical forests in a long-term drying experiment. We investigated the effects of a chronic drying manipulation on soil microbial community abundance and variation across different forests and seasons. We also compared findings with previously published data from these forests after short-term drying. This project used soils from a long-term drying experiment established in 2018 across four seasonal lowland forests in Panama. Soils were collected from 0 – 10 cm depths during three seasonal periods in control and drying plots in 2024 and 2025 from a total of 32 plots (n = 4 per forest per treatment). The forests varied in baseline rainfall and soil fertility. We calculated alpha and beta diversity indices and compared taxonomic community composition. We found significant biogeographic variation in microbial diversity and taxonomy, with significant differences across the forests and significant effects of the drying treatment. Metadata and sample IDs are within Metadata_16S.csv and Metadata_ITS.csv. Relative abundance tables of every sample at every season are shown in the Excel workbooks 16S Relative Abundance.xlsx and ITS Relative Abundance.xlsx. They are then also shown in CSV files by each taxonomic level. Linear discriminant analysis effect size (LefSe) tables are shown for the full 16S and ITS datasets (n = 96), subsets for every site at every season (n = 8), and then for the forests with each plot merged by season (n = 8). FunGuildR data table of ITS data is uploaded.

Bacteria

Data Qualification Report: SRNL Glass Composition-Properties (ComPro) Database

The Savannah River National Laboratory Glass Composition-Properties (ComPro) database is an extensive database containing pertinent composition and durability data to support the accelerated clean-up mission at the Defense Waste Processing Facility. The activities described in this data qualification report were performed to support the information contained in the database. There were two objectives of the original data qualification process. The first objective was to review supporting documentation to determine if DOE/RW-0333P Quality Assurance Requirements and Description had been implemented during the original work. If the DOE/RW-0333P Quality Assurance Requirements and Description had not been directly implemented during the original work, the second objective was to determine if the controls that were used were adequate to meet the intent of the DOE/RW-0333P Quality Assurance Requirements and Description. The results of these two objectives and the activities performed to support these decisions are described in this document. An assessment of each dataset was made to determine if the data were RW-0333P Compliant, RW-0333P Equivalent or Non-RW-0333P Compliant. The original data qualification was performed in accordance with E7, Conduct of Engineering Manual, Procedure 3.70, Revision 4, Qualification of Data. The specific method that was used was Equivalent Controls as described in E7, 3.70. Revision 2 of this document adds supporting information for the RW-0333P Compliant datasets added to Revision 3 of the database.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W

Farnsworth Unit Gas Composition Analyses

This data set includes gas composition analyses that were provided to the Southwest Regional Partnership on Carbon Sequestration (SWP) by the operators of the Farnsworth Field Unit in Ochiltree County Texas.

CO2 CCUS,Farnsworth,Gas Compositions,Morrow Sandst