Engineering PapersSearch

SEARCH · Engineering Papers

Results for “database”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

SESAME: The Los Alamos National Laboratory’s Tabulated Equation of State Database Description with Extensions for Multi-phase Representations

Modeling of the thermodynamic equation of state (EOS) of various materials has had a long storied tradition at LANL. As early as 1949 Feynman, Metropolis, and Teller published a paper presenting EOS values for some elements and a methodology for calculating the EOS at high compression [1]. Cowan and Ashkin made notable methodology improvements for compressed materials throughout the 1950’s and beyond [2]. In 1971 Jack Barnes and Jerry Rood created the SESAME database and by 1972 the database became publicly available.

36 MATERIALS SCIENCE

The National Solar Radiation Database Final Report: Fiscal Years 2022-2024

The National Solar Radiation Database (NSRDB) is the leading public source of high-resolution solar resource data in the United States, with more than 400,000 users annually. This database represents the state of the art in the satellite-based estimation of solar resource information and uses a unique physics-based modeling approach that enables improvements in accuracy with the deployment of the next-generation geostationary satellites. Making the highest-quality, state-of-the-art, regularly updated datasets available on a timely basis for users reduces costs of solar deployment by providing accurate information for siting studies and system output prediction and thereby reduces project financing costs and risks.

14 SOLAR ENERGY

Core Model Proposal #399: Updating the SSP Database (v3.0) (Population, GDP, and Labor Force) and Labor Productivity (KLEM)

This Core Model Proposal (CMP) updates the Shared Socioeconomic Pathway (SSP) database to a recent version (v3.0.1; released in 2024) within GCAM. Currently, GCAM relies on socioeconomic drivers, including population, GDP, and labor force projections, from the original SSP database version released in 2013. These projections, provided by independent socioeconomic dynamic models (e.g., multi-dimensional demographic models and macroeconomic models of convergence growth), may need regular updates when (1) near-term observations become available and (2) there are updates and advancements in the socioeconomic modeling. Timely updates of socioeconomic drivers in global economic equilibrium and multisector dynamic modeling will ensure (1) alignment of historical years and near-term projections with observations, enhancing base year calibrations, including calibration parameters and labor productivity, and (2) improvement of long-term projections with updated socioeconomic drivers, which set the scale of the economy. This CMP updates the SSP data (from v2013 to v2024) and also fixes/reconciles historical GDP data sources in GCAM. We investigate the impact of these updates on GCAM projections.

97 MATHEMATICS AND COMPUTING

Developing Source Term Database for Advanced Reactors

A source term database is crucial to informing nuclear emergency response measures, enabling emergency responders to assess the potential severity of nuclear and radiological consequences. In recent times, various advanced reactor designs have come into operation, are under construction, or are being designed and developed. This report documents an effort carried out to develop a source term database for advanced reactors. The report covers key design features of these reactors and discusses radioactivity buildup and source term inventories of dose-significant radionuclides in the reactor core. For neutronic and depletion analyses, we used the SCALE code system, a computational suite for reactor physics, depletion, criticality, and sensitivity/uncertainty quantification. We used SCALE/TRITON to perform depletion calculations to predict cycle length and discharge burnup and to generate the ORIGEN reactor library. Subsequently, we used SCALE/ORIGAMI to calculate radioactivity buildup and, thereby, the source term inventories at the targeted discharge burnup, using the ENDF/B-VII.1 nuclear data library. This report covers several advanced reactors, including the KLT-40S, RITM-200N, VOYGR, and eVinci. However, other reactors, such as the RITM-200S and ARC-100, have yet to be investigated and will be explored in future efforts.

22 GENERAL STUDIES OF NUCLEAR REACTORS

Particle Physics Division Lifting Fixture Database Restructure

The Particle Physics Division (PPD) at Fermilab needed to modernize its outdated lifting fixture database. Many fixtures lacked identification or had incomplete information. During a summer internship, I cross-referenced existing data, photographed and measured fixtures using standardized tools, and updated their locations and details in a modern Excel format. Some fixtures were identified as belonging to other departments, like the Applied Physics and Superconducting Technology Directorate (APS-TD). This effort significantly improved the accuracy and utility of the PPD database, streamlined other departmental inventories, and potentially saved hundreds of thousands of dollars. The work concluded with a clear, organized record of all verified lifting fixtures.

Creedon, Carroll [Unlisted, US, IL; Fermilab]

Model Validation Database for Fires Involving Fuels at Liquefied Natural Gas Facilities (Version 2)

This document provides a description of the model evaluation protocol (MEP) database for fires involving liquefied natural gas (LNG) and processing fuels at LNG facilities. The purpose of the MEP is to provide procedures regarding the assessment of a model’s suitability to predict thermal exclusion zones resulting from a fire. The database includes measurements from pool fire, jet fire, and fireball experiments which are provided in a spreadsheet. Users are to enter model results into the spreadsheet which automatically generates statistical performance measures and graphical comparisons with the experimental data. The intent of this document is to provide a description of the experiments and of the procedure required to carry out the validation portion of the MEP. In addition, the statistical performance measures, measurements for comparisons, and parameter variation are provided.

03 NATURAL GAS

Database Performance Monitoring for DUNE

This project improves Checkmk monitoring for DUNE Rucio PostgreSQL database services by adding clearer dashboard visibility for connection and lock behavior. The work began with a request to monitor database performance metrics such as connection usage, configured limits, lock activity, wait behavior, query performance, storage trends, and saturation alerts. Existing Checkmk PostgreSQL checks were reviewed, and gaps were identified in how connection states and lock modes were displayed. To address these gaps, two DSG-specific local checks were added for dune_rucio_prod: one for connection-state monitoring and one for lock-state monitoring. These checks report active, idle, idle-in-transaction, total, usage-percent, lock-mode, waiting-lock, and wait-age metrics. The added metrics supplement built-in Checkmk monitoring and provide DUNE application developers with clearer service states, history graphs, dashboard widgets, and alerts.

Bowers, Elliot [Cabrillo Coll.]

Fermented Foods Microbial Genomes Database

This database contains ~4,300 microbial genomes assembled from diverse fermented foods. These genomes were obtained from a larger set of 13,850 microbial genomes by clustering them at 99% average nucleotide identity (ANI) to create a "species"-representative database.

59 BASIC BIOLOGICAL SCIENCES

National Climate Database (NCDB)

The National Climate Database (NCDB) is a high resolution, bias-corrected climate dataset consisting of the three most widely used variables of solar radiation- global horizontal (GHI), direct normal (DNI), and diffuse horizontal irradiance (DHI)- as well as other meteorological data. The goal of the NCDB is to provide unbiased high temporal and spatial resolution climate data needed for renewable energy modeling. The NCDB is modeled using a statistical downscaling approach with Regional Climate Model (RCM)-based climate projections obtained from the North American Coordinated Regional Climate Downscaling Experiment (NA-CORDEX; linked below). Daily climate projections simulated by the Canadian Regional Climate Model 4 (CanRCM4) forced by the second-generation Canadian Earth System Model (CanESM2) for two Representative Concentration Pathways (RCP4.5 or moderate emissions scenario and RCP8.5 or highest baseline emission scenario) are selected as inputs to the statistical downscaling models. The National Solar Radiation Database (NSRDB) is used to build and calibrate statistical models.

Array

An Experiment with LLMs as Database Design Tutors: Persistent Equity and Fairness Challenges in Online Learning

As large language models (LLMs) continue to evolve, their capacity to replace humans as their surrogates is also improving. As increasing numbers of intelligent tutoring systems (ITSs) are embracing the integration of LLMs for digital tutoring, questions are arising as to how effective they are and if their hallucinatory behaviors diminish their perceived advantages. One critical question that is seldom asked if the availability, plurality, and relative weaknesses in the reasoning process of LLMs are contributing to the much discussed digital divide and equity and fairness in online learning. In this paper, we present an experiment with database design theory assignments and demonstrate that while their capacity to reason logically is improving, LLMs are still prone to serious errors. We demonstrate that in online learning and in the absence of a human instructor, LLMs could introduce inequity in the form of “wrongful” tutoring that could be devastatingly harmful for learners, which we call ignorant bias, in increasingly popular digital learning. We also show that significant challenges remain for STEM subjects, especially for subjects for which sound and free online tutoring systems exist. Based on the set of use cases, we formulate a possible direction for an effective ITS for online database learning classes of the future.

Jamil, Hasan M. (ORCID:0000000231243780)

CoRE MOF DB: A curated experimental metal-organic framework database with machine-learned properties for integrated material-process screening

Here, we present an updated version of the Computation-Ready, Experimental (CoRE) Metal-Organic Framework (MOF) database, which includes a curated set of computation-ready MOF crystal structures designed for high-throughput computational materials discovery. Data collection and curation procedures were improved from the previous version to enable more frequent updates in the future. Machine-learning-predicted properties, such as stability metrics and heat capacities, are included in the dataset to streamline screening activities. An updated version of MOFid was developed to provide detailed information on metal nodes, organic linkers, and topologies of an MOF structure. DDEC6 partial atomic charges of MOFs were assigned based on a machine-learning model. Gibbs ensemble Monte Carlo simulations were used to classify the hydrophobicity of MOFs. The finalized dataset was subsequently used to perform integrated material-process screening for various carbon-capture conditions using high-fidelity temperature-swing adsorption (TSA) simulations. Our workflow identified multiple MOF candidates that are predicted to outperform CALF-20 for these applications.

CoRE MOF database

Database of Nonaqueous Proton-Conducting Materials

This work presents the assembly of 48 papers, representing 74 different compounds and blends, into a machine-readable database of nonaqueous proton-conducting materials. SMILES was used to encode the chemical structures of the molecules, and we tabulated the reported proton conductivity, proton diffusion coefficient, and material composition for a total of 3152 data points. The data spans a broad range of temperatures ranging from -70 to 260 °C. To explore this landscape of nonaqueous proton conductors, DFT was used to calculate the proton affinity of 18 unique proton carriers. The results were then compared to the activation energy derived from fitting experimental data to the Arrhenius equation. It was found that while the widely recognized positive correlation between the activation energy and proton affinity may hold among closely related molecules, this correlation does not necessarily apply across a broader range of molecules. This work serves as an example of the potential analyses that can be conducted using literature data combined with emerging research tools in computation and data science to address specific materials design problems.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Metallic fuel transient fuel-cladding interface liquefaction model assessment platform enabled by integrating BISON with databases

A novel platform has been developed within the BISON fuel performance code to assess models of fuel-cladding interface liquefaction for sodium-cooled fast reactor (SFR) metallic fuels. Here, this platform is crucial because liquefaction at the fuel-cladding interface significantly impacts fuel performance and may compromise fuel pin integrity during transient events. To ensure accurate predictions, the platform integrates data collected during the Integral Fast Reactor (IFR) program, now archived in metallic fuel databases. This integration supports verification and validation (V&V) of the models in BISON. Leveraging the extensive US experience with metallic fuel liquefaction and the collections of preserved legacy data, the platform serves as a powerful tool for evaluating existing models and advancing the development of new ones.

11 - NUCLEAR FUEL CYCLE AND FUEL MATERIALS

Autogenerating a Domain-Specific Question-Answering Data Set from a Thermoelectric Materials Database to Enable High-Performing BERT Models

We present a method for autogenerating a large domain-specific question-answering (QA) dataset from a thermoelectric materials database. We show that a small language model, BERT, once fine-tuned on this automatically generated dataset of 99,757 QA pairs about thermoelectric materials, affords better performance in the field of thermoelectric materials compared to a BERT model fine-tuned on the generic English-language QA data set, SQuAD-v2. We further show that mixing the two data sets (ours and SQuAD-v2), which have significantly different syntactic and semantic scopes, allows the BERT model to achieve even better performance. The best-performing BERT model fine-tuned on the mixed data set outperforms the models fine-tuned on the other two data sets by scoring an exact match of 67.93% and an F1 score of 72.29% when evaluated on our test data set. This has important implications as it demonstrates the ability to realize high-performing small language models, with modest computational resources, empowered by domain-specific materials data sets which can be generated according to our method.

biological databases

JINAbase: A database for chemical abundances of metal-poor stars

CeNAM maintains the Stellar Abundance Database JINAbase provides detailed abundance information of 2766 stars from 173 publications. The user interface enables easy graphing of user selected element ratios to explore trends and scatter.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS

Scintillator Library: A database of inorganic and organic scintillator properties

The Scintillator Library (scintillator.lbl.gov) is a database of scintillator properties hosted by Lawrence Berkeley National Laboratory in a web-accessible format. It contains a variety of measured inorganic and organic scintillator properties extracted from peer-reviewed literature and manufacturer specifications. Data housed within the Scintillator Library supply an important resource for developers of scintillator-based detection systems and an aid for scientists seeking to establish connections between fundamental material and chemical properties and the associated scintillation performance.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND

CHUWD-H v1.0: a comprehensive historical hourly weather database for U.S. urban energy system modeling

Reliable and continuous meteorological data are crucial for modeling the responses of energy systems and their components to weather and climate conditions, particularly in densely populated urban areas. However, existing long-term datasets often suffer from spatial and temporal gaps and inconsistencies, posing great challenges for detailed urban energy system modeling and cross-city comparison under realistic weather conditions. Here we introduce the Historical Comprehensive Hourly Urban Weather Database (CHUWD-H) v1.0, a 23-year (1998-2020) gap-free and quality-controlled hourly weather dataset covering 550 weather station locations across all urban areas in the contiguous United States. CHUWD-H v1.0 synthesizes hourly weather observations from stations with outputs from a physics-based solar radiation model and a reanalysis dataset through a multi-step gap filling approach. A 10-fold Monte Carlo cross-validation suggests that the accuracy of this gap filling approach surpasses that of conventional gap filling methods. Designed primarily for urban energy system modeling, CHUWD-H v1.0 should also support historical urban meteorological and climate studies, including the validation and evaluation of urban climate modeling.

54 ENVIRONMENTAL SCIENCES

PAVC: The foundation for a Pan-Arctic Vegetation Cover database

Field-measured Arctic vegetation cover data is essential for creating accurate, high-quality vegetation structure and composition maps. Extrapolating field data into high-resolution cover maps provides detailed, function-specific information for use in Earth System Models, vegetation classifications, and monitoring vegetation change over time and space. However, field campaigns that collect plant cover vary substantially in scope, method, and purpose, which makes them difficult to unify across data stores, and they are often not designed to meet remote sensing needs. In this work, we synthesized and harmonized field-based fractional cover data from various data stores to create a high-quality, consistent repository schema for remote sensing-based vegetation cover mapping applications. We developed a reproducible workflow for synthesizing visual estimate and point-intercept fractional cover data. The resultant Pan-Arctic Vegetation Cover (PAVC) database contains synthesized fractional cover at both the species and plant functional type levels. The latter includes absolute foliar cover for deciduous shrubs and trees, evergreen shrubs and trees, forbs, graminoids, lichen, bryophytes, and “other” vegetation, as well as absolute cover for litter and top cover for water and bare ground.

Steckler, Morgan R. [Oak Ridge National Laboratory