Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “databases”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 181 records · Page 10

Linking Extragalactic Transients and Their Host Galaxy Properties: Transient Sample, Multiwavelength Host Identification, and Database Construction

Understanding the preferences of transient types for host galaxies with certain characteristics is key to studies of transient physics and galaxy evolution, as well as to transient identification and classification in the LSST era. Here we describe a value-added database of extragalactic transients—supernovae, tidal disruption events, gamma-ray bursts, and other rare events—and their host galaxy properties. Based on reported coordinates, redshifts, and host galaxies (if known) of events, we cross-identify their host galaxies or most likely host candidates in various value-added or survey catalogs, and compile the existing photometric, spectroscopic, and derived physical properties of the host galaxies in these catalogs. This new database covers photometric measurements from the far-ultraviolet to mid-infrared. Spectroscopic measurements and derived physical properties are also available for a smaller subset of hosts. For our 36,333 unique events, we have cross-identified 13,753 host galaxies using host names, plus 4480 using host coordinates. Besides those with known hosts, there are 18,100 transients with newly identified host candidates. This large database will allow explorations of the connections of transients to their hosts, including a path toward transient alert filtering and probabilistic classification based on host properties.

79 ASTRONOMY AND ASTROPHYSICS↗

The CAI Database: 26 Al– 26 Mg Isotope Systematics

We present a publicly available calcium–aluminum-rich inclusion (CAI) database that focuses on the initial 26 Al/ 27 Al 0 ratio in CAIs, designed in a way that researchers in cosmochemistry and astrophysics may find useful. To date, the database contains 497 CAIs from 75 peer-reviewed papers. The CAIs are from all chondrite groups and cover different CAI types, textures, and sizes. The database includes the paper; the host meteorite; the CAI name and type; the 26 Al/ 27 Al 0 , δ 26 Mg$^*_0$, and δ 25 Mg values and their uncertainties; the number of regression points; the maximum 27 Al/ 24 Mg; the mean-squared weighted deviation; the CAI size; and CAI descriptions. We grouped the CAIs in different ways to discuss 26 Al/ 27 Al 0 ratio distributions with implications for the CAI formation timeline. Overall, we agree with previous authors that CAIs have a bimodal 26 Al distribution: CAIs with robust isochrons (n = 151) have a median 26 Al/ 27 Al 0 = 4.8 × 10 −5 (with a 1σ standard error of 0.1), while those with isotopic anomalies (n = 87) have a median 26 Al/ 27 Al 0 = 0.3 × 10 −5 (with a 1σ standard error of 0.2). However, the large standard deviation of both groups (1.3 and 2.3, respectively) indicates that the 26 Al/ 27 Al 0 values scatter significantly within each population. CAI types and groups can have distinct 26 Al/ 27 Al 0 and δ 26 Mg$^*_0$, but the unmelted inclusions (n = 33) have the highest median 26 Al/ 27 Al 0 = 5.1 × 10 −5 and a low median δ 26 Mg$^*_0$ = −0.05‰. We find slightly different 26 Al/ 27 Al 0 distributions between CAI chondrite types, but no differences between petrographic types or sizes. These observations can help us to understand CAI formation in the context of astrophysical models.

Astronomy and AstroPhysics↗

SoDaH: the SOils DAta Harmonization database, an open-source synthesis of soil data from research networks, version 1.0

Data collected from research networks present opportunities to test theories and develop models about factors responsible for the long-term persistence and vulnerability of soil organic matter (SOM). Synthesizing datasets collected by different research networks presents opportunities to expand the ecological gradients and scientific breadth of information available for inquiry. Synthesizing these data is challenging, especially considering the legacy of soil data that have already been collected and an expansion of new network science initiatives. To facilitate this effort, here we present the SOils DAta Harmonization database (SoDaH; https://lter.github.io/som-website, last access: 22 December 2020), a flexible database designed to harmonize diverse SOM datasets from multiple research networks. SoDaH is built on several network science efforts in the United States, but the tools built for SoDaH aim to provide an open-access resource to facilitate synthesis of soil carbon data. Moreover, SoDaH allows for individual locations to contribute results from experimental manipulations, repeated measurements from long-term studies, and local- to regional-scale gradients across ecosystems or landscapes. Finally, we also provide data visualization and analysis tools that can be used to query and analyze the aggregated database. The SoDaH v1.0 dataset is archived and available at https://doi.org/10.6073/pasta/9733f6b6d2ffd12bf126dc36a763e0b4 (Wieder et al., 2020).

54 ENVIRONMENTAL SCIENCES↗

The ABCflux database: Arctic–boreal CO 2 flux observations and ancillary information aggregated to monthly time steps across terrestrial ecosystems

Past efforts to synthesize and quantify the magnitude and change in carbon dioxide (CO 2 ) fluxes in terrestrial ecosystems across the rapidly warming Arctic–boreal zone (ABZ) have provided valuable information but were limited in their geographical and temporal coverage. Furthermore, these efforts have been based on data aggregated over varying time periods, often with only minimal site ancillary data, thus limiting their potential to be used in large-scale carbon budget assessments. To bridge these gaps, we developed a standardized monthly database of Arctic–boreal CO 2 fluxes (ABCflux) that aggregates in situ measurements of terrestrial net ecosystem CO 2 exchange and its derived partitioned component fluxes: gross primary productivity and ecosystem respiration. The data span from 1989 to 2020 with over 70 supporting variables that describe key site conditions (e.g., vegetation and disturbance type), micrometeorological and environmental measurements (e.g., air and soil temperatures), and flux measurement techniques. Here, we describe these variables, the spatial and temporal distribution of observations, the main strengths and limitations of the database, and the potential research opportunities it enables. In total, ABCflux includes 244 sites and 6309 monthly observations; 136 sites and 2217 monthly observations represent tundra, and 108 sites and 4092 observations represent the boreal biome. The database includes fluxes estimated with chamber (19 % of the monthly observations), snow diffusion (3 %) and eddy covariance (78 %) techniques. The largest number of observations were collected during the climatological summer (June–August; 32 %), and fewer observations were available for autumn (September–October; 25 %), winter (December–February; 18 %), and spring (March–May; 25 %). ABCflux can be used in a wide array of empirical, remote sensing and modeling studies to improve understanding of the regional and temporal variability in CO 2 fluxes and to better estimate the terrestrial ABZ CO 2 budget.

59 BASIC BIOLOGICAL SCIENCES↗

JHTDB-wind: a web-accessible large-eddy simulation database of a wind farm with virtual sensor querying

This paper introduces JHTDB-wind (https://turbulence.idies.jhu.edu/datasets/windfarms, last access: 11 November 2025), a publicly accessible database containing large-eddy simulation (LES) data from wind farms. Building on the framework of the Johns Hopkins Turbulence Database (JHTDB), which hosts direct numerical simulation (DNS) and some LES datasets of canonical turbulent flows, JHTDB-wind stores the 4D space–time history of the flow and provides users the ability to access and query the data via a web-based virtual sensor interface. The initial dataset comprises LES results from a large wind farm with 10×6 turbines, modeled using a filtered actuator line method, under conventionally neutral atmospheric conditions. These data comprise 1 h (hour) of flow field data (velocity, pressure, potential temperature deviation, subgrid-scale (SGS) eddy viscosity, and turbine forces, approximately 15 TB (terabytes) and wind turbine data – including both turbine-level operational quantities and blade-level aerodynamic quantities (approximately 1.3 TB) – stored in Zarr and Parquet formats, respectively. Data retrieval is facilitated by the giverny Python package, allowing remote users to query the database in Python or MATLAB (C and Fortran support are available for flow field data). This paper details the simulation setup and demonstrates data access through examples that analyze wind farm flow structures and turbine performance. The framework is extensible to future datasets, including the JHTDB-wind diurnal cycle simulation analyzed in Xiao et al. (2025).

17 WIND ENERGY↗

ECAR-1930 BASELINE CHARACTERIZATION DATABASE VERIFICATION REPORT - NBG-18 BILLET 635-14

The purpose of this Engineering Calculations and Analysis Report is to present the data being collected in the Baseline Graphite Characterization program, which is directly tasked with supporting the Idaho National Laboratory’s (INL’s) research and development efforts on the Next Generation Nuclear Plant (NGNP)/Very High Temperature Reactor (VHTR). This program is populating a comprehensive database that will reflect the baseline properties of nuclear-grade graphite with regard to individual grade, billet, and position within individual billets. The physical and mechanical property information being collected will be transferred to the NGNP Data Management and Analysis System (NDMAS), and from that database will help populate handbook of property data available to member nations of the Generation IV International Forum (GIF). The transfer of this data from the applicable technical lead to the dissemination databases available to other end users requires a full review of the test procedures and data collection efforts through an analysis of the multiple summary spreadsheets and values being collected. This report represents that analysis for a single billet of nuclear grade graphite (NBG-18 billet 635-14) and facilitates the release of the associated data to the NDMAS custodians.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Baseline Characterization Database Verification Report - PCEA Billet 02S8-5

The purpose of this engineering calculations and analysis report (ECAR) is to present data collected in the Baseline Graphite Characterization Program, which is directly tasked with supporting the Idaho National Laboratory's (INL's) research and development efforts on the Advanced Reactor Technologies (ART) Program. This program populates a comprehensive database that reflects the baseline properties of nuclear-grade graphite regarding individual grade, billet, and position within individual billets. The physical- and mechanical-property information being collected will be transferred to the Nuclear Data Management and Analysis System (NDMAS), and that database will help populate the handbook of property data available to member nations of the Generation-IV International Forum. Transfer of these data from the applicable technical lead to the dissemination databases available to other end users requires a full review of the test procedures and data-collection efforts through an analysis of the multiple summary spreadsheets and values being collected. This report represents the analysis for PCEA Billet 02S8-5 and facilitates release of associated data to the NDMAS custodians.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Baseline Characterization Database Verification Report - PCEA Billet 01S8-9

The purpose of this engineering calculations and analysis report (ECAR) is to present data collected in the Baseline Graphite Characterization Program, which is directly tasked with supporting the Idaho National Laboratory's (INL's) research and development efforts on the Advanced Reactor Technologies (ART) Program. This program populates a comprehensive database that reflects the baseline properties of nuclear-grade graphite with regard to individual grade, billet, and position within individual billets. The physical- and mechanical-property information being collected will be transferred to the Nuclear Data Management and Analysis System (NDMAS), and that database will help populate the handbook of property data available to member nations of the Generation-IV International Forum. Transfer of these data from the applicable technical lead to the dissemination databases available to other end users requires a full review of the test procedures and data-collection efforts through an analysis of the multiple summary spreadsheets and values being collected. This report represents the analysis for PCEA Billet 01S8-9 and facilitates release of associated data to the NDMAS custodians.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Baseline Characterization Database Verification Report ? NBG-17 Billet V104

The purpose of this report is to present data collected in the Baseline Graphite Characterization Program, which is directly tasked with supporting the Idaho National Laboratory’s (INL’s) research and development efforts on the Advanced Reactor Technologies (ART) Program. This program populates a comprehensive database that reflects the baseline properties of nuclear-grade graphite regarding individual grade, billet, and position within individual billets. The physical- and mechanical-property information being collected will be transferred to the Nuclear Data Management and Analysis System (NDMAS), and that database will help populate the handbook of property data available to member nations of the Generation-IV International Forum. Transfer of these data from the applicable technical lead to the dissemination databases available to other end users requires a full review of the test procedures and data-collection efforts through an analysis of the multiple summary spreadsheets and values being collected. This report represents the analysis for NBG-17 Billet V104 and facilitates release of associated data to the NDMAS custodians. Millions of raw data points have been collected during testing and quantification analyses for these billets. The summary scalar property values and supplementary traceability data are collected into comprehensive spreadsheets. Data sets are composed of single billets of graphite for any given grade, organized by mechanical test-specimen type, and further subdivided into individual spreadsheet tabs according to the specific test or evaluation being performed. A direct analysis of properties was not conducted, and this report does not provide information on the validity or performance characteristics of the graphite itself. Rather, this report is intended as a verification of the completeness of actual data collected in accordance with PLN-3467, “Baseline Graphite Characterization Plan: Electromechanical Testing,” [1] and PLN-3348 “Graphite Mechanical Testing” [2] and their representation of the measurement and test results with sole regard to the graphite billets under evaluation.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Populating the Hydrogen Component Reliability Database (HYCRED) with Incident Data from Hydrogen Dispensing

Safety, risk, and reliability issues are vital to ensure the continuous and profitable operation of hydrogen technologies. Quantitative risk assessment (QRA) has been used to enable the safe deployment of engineering systems, especially hydrogen fueling stations. However, QRA studies require reliability data which are essential to collect to make the studies as realistic and relevant as possible. These data are currently lacking and data from other industries, such as oil and gas, are used in hydrogen system QRAs. This may lead to inaccurate results since hydrogen fueling stations have differences in physical properties, system design, and operational parameters when compared to other fueling stations, thus necessitating new data sources are necessary to capture the effects of these differences. To address this gap, we developed a structure for a hydrogen component reliability database, (HyCReD) [1], which could be used to generate reliability data to be used in QRA studies. In this paper, we demonstrate populating the HyCReD database with information extracted from new narrative reports on hydrogen fueling station incidents, specifically focused on the dispensing processes. We analyze five new events and demonstrate the feasibility of populating the database and types of meaningful insights that can be obtained at this stage.

component reliability↗

SG50 Data-format Specifications Document for the Automatically Readable, Comprehensive, and Curated Experimental Reaction Database MEDUSAL

The aim of this document is to lay out a first draft of the specifications for the MEDUSAL database (Machine-readable Experimental Data User App & Library) that is being described by OECD/NEA/WPEC SG-50. The EXFOR database (Otuka et al., 2014) has a format that is based on code-value pairs, and a significant portion of the information in the EXFOR entry is contained in free text sections. Several high-level requirements for the MEDUSAL database, as laid out in the Use Cases and Requirements Working Paper (citation), relate to the definition of the specifications

Nuclear Criticality Safety Program (NCSP)↗

Database-Agnostic Log Analysis and Monitoring Framework

Prior to my internship, I was informed that a previous intern had built a tool to analyse MongoDB logs and look for invalid access attempts, which served as a great reference point for my project. I was initially tasked with expanding on her prototype and filling in the gaps such as integrating it with the main monitoring tool the lab uses. Eventually, the scope grew, expanding to support other databases and a growing collection of tools. I organized the framework around an observer pattern, meaning one point in the program sending updates to the rest of the framework. Every time a log was read and parsed, it was sent to be processed by the tools, using the type of event as a means to determine which tools should get a chance to act on the log. This decouples the tools from the log reader, making future updates and additions much easier. The framework processes MongoDB logs at ~135,000 entries per second and PostgreSQL logs at ~170,500 entries per second, accurately detecting anomalies such as slow queries and connections from unknown addresses. This framework serves to fill gaps in database monitoring tools currently implemented at the lab, such as tracking failed authentication for PostgreSQL and MongoDB which had very minimal or none before this framework. National labs such as Fermilab hold sensitive data and valuable computing resources, making them attractive targets. Monitoring intrusion attempts on databases is made much easier by this comprehensive monitoring suite.

Clark, Dylan [Unlisted, US, IL; Fermilab]↗

Optimized thermodynamic properties of REE aqueous species (REE 3+ and REEOH 2+ ) and experimental database for modeling the solubility of REE phosphate minerals (monazite, xenotime, and rhabdophane) from 25 to 300 °C

Rare earth elements (REE) are critical elements found in monazite, xenotime, and hydrated REE phosphates which typically form in hydrothermal mineral deposits. Accurate predictions of the solubility of these REE phosphates and the speciation of REE in aqueous fluids are both key to understanding the controls on the transport, fractionation, and deposition of REE in natural systems. Previous monazite and xenotime solubility experiments indicate the presence of large discrepancies between experimentally derived solubility constants versus calculated solubilities by combining different data sources for the thermodynamic properties of minerals and aqueous species at hydrothermal conditions. In this study, these discrepancies were resolved by using the program GEMSFITS to optimize the standard partial molal Gibbs energy of formation (Δ f G° 298 ) of REE aqueous species (REE 3+ and REE hydroxyl complexes) at 298.15 K and 1 bar while keeping the thermodynamic properties fixed for the REE phosphates. A comprehensive experimental database was compiled using solubility data available between 25 and 300 °C. The latter permits conducting thermodynamic parameter optimization of Δ f G° 298 for REE aqueous species. Optimal matching of the rhabdophane solubility data between 25 and 100 °C requires modifying the Δ f G° 298 values of REE 3+ by 1–6 kJ/mol, whereas matching of the monazite solubility data between 100 and 300 °C requires modifying the Δ f G° 298 values of both REE 3+ and REEOH 2+ by ~15–31 kJ/mol and ~2–10 kJ/mol, respectively. For xenotime, adjustments of Δ f G° 298 values by 1–26 kJ/mol are only necessary for the REE 3+ species. The optimizations indicate that the solubility of monazite in acidic solutions is controlled by the light (L)REE 3+ species at <150 °C and the LREEOH 2+ species at >150 °C, whereas the solubility of xenotime is controlled by the heavy (H)REE 3+ species between 25 and 300 °C. Based on the optimization results, we conclude that the revised Helgeson-Kirkham-Flowers equation of state does not reliably predict the thermodynamic properties of REE 3+ , REEOH 2+ , and likely other REE hydroxyl species at hydrothermal conditions. We therefore provide an experimental database (ThermoExp_REE) as a basic framework for future updates, extensions with other ligands, and optimizations as new experimental REE data become available. As a result, the optimized thermodynamic properties of aqueous species and minerals are available open access to accurately predict the solubility of REE phosphates in fluid-rock systems.

58 GEOSCIENCES↗

Information-Theoretic Exploration of Multivariate Time-Varying Image Databases

Modern scientific simulations produce very large datasets, making interactive exploration of such data computationally prohibitive. An increasingly common data reduction technique is to store visualizations and other data extracts in a database. The Cinema project is one such approach, storing visualizations in an image database for post hoc exploration and interactive image-based analysis. This work focuses on developing efficient algorithms that can quantify various types of multivariate dependencies existing within multi-variable datasets. It applies specific mutual information measures for the quantification of salient regions from multivariate image data. Here, using such information measures, the opacity of the images is modulated so that the salient regions are automatically highlighted and the domain scientists can interactively explore the most relevant regions for scientific discovery.

97 MATHEMATICS AND COMPUTING↗

Generation of macro- and microplastic databases by high-throughput FTIR analysis with microplate readers

Abstract FTIR spectral identification is today’s gold standard analytical procedure for plastic pollution material characterization. High-throughput FTIR techniques have been advanced for small microplastics (10–500 µm) but less so for large microplastics (500–5 mm) and macroplastics (> 5 mm). These larger plastics are typically analyzed using ATR, which is highly manual and can sometimes destroy particles of interest. Furthermore, spectral libraries are often inadequate due to the limited variety of reference materials and spectral collection modes, resulting from expensive spectral data collection. We advance a new high-throughput technique to remedy these problems using FTIR microplate readers for measuring large particles (> 500 µm). We created a new reference database of over 6000 spectra for transmission, ATR, and reflection spectral collection modes with over 600 plastic, organic, and mineral reference materials relevant to plastic pollution research. We also streamline future analysis in microplate readers by creating a new particle holder for transmission measurements using off-the-shelf parts and fabricating a nonplastic 96-well microplate for storing particles. We determined that particles should be presented to microplate readers as thin as possible due to thick particles causing poor-quality spectra and identifications. We validated the new database using Open Specy and demonstrated that additional transmission and reflection spectra reference data were needed in spectral libraries. Graphical abstract

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

The Intermetallic Reactivity Database: Compiling Chemical Pressure and Electronic Metrics toward Materials Design and Discovery

Here, the advent of high-throughput Density Functional Theory (DFT) calculations has supported the creation of large databases containing the quantitative output necessary for constructing theoretical phase diagrams and predicting physical properties. In this Article, we present a complementary resource, the Intermetallic Reactivity Database (IRD), focused on the chemical bonding features of solid-state structures and indicators of potential structural transformations. Each IRD entry augments common features, such as band structures and density of states (DOS) distributions, with chemically motivated information including DFT-Chemical Pressure (CP) schemes and visualizable representations of the atomic charges. Together, these data types enable the rationalization and prediction of potential structural phenomena encountered in intermetallic chemistry, as we illustrate with four examples: the origins of the Y 2 Ni 2 Mg structure in terms of CP features of its parent structures, the anticipation of intergrowth phases from the net atomic CPs collected in Al-containing binary phases, the correlation between trends in the CP schemes of CaCu 5 -type phases and experimentally observed structural variations, and finally, the development of theoretical methodology with the testing of a streamlined method generating DFT-CP schemes. Altogether, these examples highlight how the IRD supports the creation of models of structural chemistry that extend beyond the bounds of its entries.

36 MATERIALS SCIENCE↗

Plant Metabolic Network 15: A resource of genome-wide metabolism databases for 126 plants and algae

To understand and engineer plant metabolism, we need a comprehensive and accurate annotation of all metabolic information across plant species. As a step towards this goal, we, in this study, generated genome-scale metabolic pathway databases of 126 algal and plant genomes, ranging from model organisms to crops to medicinal plants (https://plantcyc.org). Of these, 104 have not been reported before. We systematically evaluated the quality of the databases, which revealed that our semi-automated validation pipeline dramatically improves the quality. We then compared the metabolic content across the 126 organisms using multiple correspondence analysis and found that Brassicaceae, Poaceae, and Chlorophyta appeared as metabolically distinct groups. To demonstrate the utility of this resource, we used recently published sorghum transcriptomics data to discover previously unreported trends of metabolism underlying drought tolerance. We also used single-cell transcriptomics data from the Arabidopsis root to infer cell type-specific metabolic pathways. This work shows the quality and quantity of our resource and demonstrates its wide-ranging utility in integrating metabolism with other areas of plant biology.

59 BASIC BIOLOGICAL SCIENCES↗

Compilation of a Solar Mirror Materials Database and an Analysis of Natural and Accelerated Mirror Exposure and Degradation

The National Renewable Energy Laboratory (NREL) has been conducting exposure experiments on solar reflectors for over four decades. Thousands of mirror samples from over one hundred suppliers have been exposed to and monitored in a range of relevant environments. These test conditions include outdoor test settings and several controlled laboratory environments. These samples have been rigorously individually characterized using a series of reflectance measurements, visual inspections, and in some cases, in-depth composition analysis to identify degradation modes, reflectance losses, and other mirror properties integral to understanding the solar reflector's life cycle. Here, this paper compiles the decades of measurement data into a concise statistical analysis. It includes exposure and degradation data for numerous reflector types, including secondary-surface reflector permutations of polymer and glass superstrates with silver and aluminum reflectors as well as front-surface reflectors. The results herein are intended to analyze environmental stressors and degradation trends among various historical and state-of-the-art solar reflectors. It may be used to support solar reflector design, effective testing methodology, and inform manufacturing decisions moving forward. Presented are the results of the compiled database and an initial analysis for degradation rate modeling using full-spectrum and wavelength-dependent approaches. The database is a growing resource hosted on a live, publicly accessible website. In conjunction with the analysis presented here, it provides a valuable resource to the solar reflector manufacturing and testing industry.

14 SOLAR ENERGY↗