Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “databases”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 307 records · Page 17

CONUS and Global Carbon Consumption Database for Wildland Fire

Fire plays a significant role on both national and global scales, profoundly impacting landscapes shaped by human activity as well as wildlands. Even though fire can be devastating, wildland fire is a natural and integral force on our landscapes. Fires can also serve to reduce fuels to mitigate wildfire risk and maintain healthy ecosystem functions. However, the smoke produced by fires, regardless of their size or purpose, can pose adverse effects on human health when inhaled downwind. Understanding the influence of smoke on air quality and human well-being necessitates the quantification of emissions that fires release into the atmosphere. In response to this need, we have developed a carbon consumption database for the Continental United States (CONUS) and are developing a global carbon consumption database, both of which are directly link to distinct fuels within various fire danger categories. Fuel Characteristic Classification System (FCCS) fuels are used to parameterize biomass, and consumption is broken down into five Fire Danger categories (Low, Moderate, High, Very High, Extreme), for both ‘new’ and ‘residual’ burning scenarios. Residual burned area is defined as burning in areas that have recently burned. We implemented the FCCS 30-meter CONUS fuelbed dataset during the 2019 Fire Influence on Regional to Global Environments and Air Quality (FIREX-AQ) campaign to estimate daily carbon emissions. Our emissions estimates were rigorously compared against in-situ measurements of CO2, CO, and black carbon aerosols, revealing a robust agreement between the two datasets. The 300-meter global database builds upon the foundations of the Pettinari, M. Lucrecia (2015) Global Fuelbed database, a global fuel map with standardized FCCS biomass parameters. These products can serve as valuable tools when used in conjunction with burned area data to rapidly and accurately estimate carbon consumed and released into the atmosphere.

FCCS↗

The CCS-EJ-SJ Database: A Tool for Addressing Justice Challenges in Carbon Capture and Storage

At CCUS 2024, within Theme 8: ESG and Stakeholder Engagement, the poster presentation on "The CCS-EJ-SJ Database: A Tool for Addressing Justice Challenges in Carbon Capture and Storage" highlights the innovative database designed to integrate environmental, social, energy, and economic justice into CCS projects. The CCS-EJ-SJ Database serves as a comprehensive resource for stakeholders, offering an interactive dashboard that provides access to essential data across various justice themes. It incorporates established initiatives like the Justice40 Initiative and tools such as the DOE's Disadvantaged Community Reporter, enhancing its utility and credibility. This presentation will highlight the CCS-EJ-SJ database's framework, the consequential research products, and its pivotal role in offering crucial insights into CCS opportunities, ultimately expediting sustainable, environmentally considerate, and socially responsible efforts.

White, Casey↗

Database Performance Monitoring for DUNE

This report presents the research, design, and implementation of improved PostgreSQL monitoring for DUNE Rucio database services using Checkmk. The project began with a request to improve dashboard visibility for database performance metrics, including connection usage, configured connection limits, lock activity, wait behavior, storage trends, query performance, and saturation alerts. The initial implementation focused on the dune_rucio_prod database on the rucio_prod PostgreSQL instance because connection saturation and lock contention are direct reliability risks for database-backed services. Existing Checkmk PostgreSQL monitoring was investigated, and several gaps were identified. Built-in connection monitoring did not clearly separate active, idle, idle-in-transaction, total, and usage-percent metrics, while the built-in lock monitoring simplified PostgreSQL lock modes into shared and exclusive categories. To address these gaps, two DSG-specific Checkmk local checks were created: one for connection-state monitoring and one for lock-state monitoring. These checks supplement the built-in PostgreSQL checks and provide additional performance data for dashboard graphs, service states, and alerts.

Bowers, Elliot [Cabrillo Coll.]↗

The IsoGenie database: an interdisciplinary data management solution for ecosystems biology and environmental research

Modern microbial and ecosystem sciences require diverse interdisciplinary teams that are often challenged in “speaking” to one another due to different languages and data product types. Here we introduce the IsoGenie Database, a de novo developed data management and exploration platform, as a solution to this challenge of accurately representing and integrating heterogenous environmental and microbial data across ecosystem scales. The IsoGenieDB is a public and private data infrastructure designed to store and query data generated by the IsoGenie Project, a ~10 year DOE-funded project focused on discovering ecosystem climate feedbacks in a thawing permafrost landscape. The IsoGenieDB provides (i) a platform for IsoGenie Project members to explore the project’s interdisciplinary datasets across scales through the inherent relationships among data entities, (ii) a framework to consolidate and harmonize the datasets needed by the team’s modelers, and (iii) a public venue that leverages the same spatially explicit, disciplinarily integrated data structure to share published datasets. The IsoGenieDB is also being expanded to cover the NASA-funded Archaea to Atmosphere (A2A) project, which scales the findings of IsoGenie to a broader suite of Arctic peatlands, via the umbrella A2A Database (A2A-DB). The IsoGenieDB’s expandability and flexible architecture allow it to serve as an example ecosystems database.

54 ENVIRONMENTAL SCIENCES↗

A comprehensive diffusion mobility database comprising 23 elements for magnesium alloys

We report that reliable experimental diffusion coefficients of 10 key alloying elements in Mg obtained by the present authors together with experimental data in the literature enabled us to perform a systematic test of the reliability of diffusion coefficients obtained from DFT calculations. The computed activation energy values were found to be quite accurate (mostly within 0.2 eV) but the computed pre-factors were less reliable. Such insights allowed us to develop a practical and yet robust strategy to perform diffusion mobility assessments by adopting the computed activation energy while fitting only the pre-factor when available experimental data are limited to a narrow temperature range. The overall good agreement between the DFT data and experimental data also gave us the confidence to employ the computed data for those that were still missing or inaccessible from experimental measurements. A systematic assessment of both the measured and computed diffusion data in hcp Mg was performed using the above holistic approach to yield the most comprehensive open Mg mobility database to date, comprising 23 elements (Mg, Ag, Al, Be, Ca, Cd, Ce, Cu, Fe, Ga, Gd, In, La, Li, Mn, Nd, Ni, Pu, Sb, Sn, U, Y, Zn). This more reliable mobility database will contribute to future development of advanced Mg alloys. The holistic approach developed in this study will be very beneficial to the future establishment of reliable mobility databases for other alloy systems as well.

36 MATERIALS SCIENCE↗

A new database website for nuclear level densities

We introduce a new open-access, web-based database (http://nld.ascsn.net), Current Archive of Nuclear Density of Levels (CANDL), that hosts experimental nuclear level density (NLD) datasets from a variety of techniques and energy ranges. Built using the Dash framework in Python, the database is designed to be interactive and user-friendly, allowing researchers to search, visualize, fit, and export NLD data with minimal effort. This resource includes data extracted from evaporation spectra, Oslo method variants, and other experimental techniques that cover excitation energies beyond the neutron resonance region. The database supports on-the-fly fitting with two widely-used phenomenological models—the Constant Temperature (CT) model and the Back-Shifted Fermi Gas (BSFG) model—selected for their simplicity and computational efficiency. Future versions aim to include additional datasets and model types, as well as easy-to-use interfaces to data science techniques. Here, this platform offers a vital tool for the nuclear physics, astrophysics, medicine, and reactor design communities.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

CatTestHub: A benchmarking database of experimental heterogeneous catalysis for evaluating advanced materials

The ability to quantitatively compare newly evolving catalytic materials and technologies is hindered by the widespread availability of catalytic data collected in a consistent manner. While certain catalytic chemistries have been widely studied across decades of scientific research, quantitative comparisons based on literature information is hindered by variability in reaction conditions, types of reported data, and reporting procedures. Here, we present CatTestHub, an open-access database dedicated to benchmarking experimental heterogeneous catalysis data. Combining systematically reported catalytic activity data for selected probe chemistries, with relevant material characterization and reactor configuration information, the database provides a collection of catalytic benchmarks for distinct classes of active site functionality. Through key choices in data access, availability, and traceability, CatTestHub seeks to balance the fundamental information needs of chemical catalysis and the FAIR data design principles. Details of the database architecture and the means through which to navigate it are presented, highlighting examples of catalytic insights readily drawn from the available benchmarking data. In its current iteration, CatTestHub spans over 250 unique experimental data points, collected over 24 solid catalysts, that facilitated the turnover of 3 distinct catalytic chemistries. Here, a roadmap is presented through which to expand the open-access platform that serves as a community wide benchmark, primarily through continuous addition of kinetic information on select catalytic systems by members of the heterogeneous catalysis community at large.

Benchmark↗

The Baghdad Atlas: A relational database of inelastic neutron-scattering (n,n ' γ) data

A relational database has been developed based on the original (n,n'γ) work carried out by A. M. Demidov et al., at the Nuclear Research Institute in Baghdad, Iraq (Demidov et al., 1978) for 105 independent measurements comprising 76 elemental samples of natural composition and 29 isotopically-enriched samples. The information from this Atlas includes: γ-ray energies and relative intensities; nuclide and level data corresponding to the residual nucleus and meta data associated with the target sample that allows for the extraction of the flux-weighted (n,n'γ) cross sections for a given transition relative to a defined value. The optimized angular-distribution-corrected fast-neutron flux-weighted partial γ-ray cross section for the production of the 846.8-keV 21+→0gs+γ-ray transition in 56Fe, determined to be $\langle$σγ$\rangle$=143(29) mb, is used for this purpose. However, different values for the adopted cross section can be readily implemented to accommodate user preference based on revised determinations of this quantity. The Atlas (n,n'γ) data has been compiled into a series of CSV-style ASCII data sets and a suite of Python scripts have been developed to build and install the database locally. The database can then be accessed directly through the SQLite engine, or using alternative methods such as the Jupyter Notebook Python-browser interface. Several examples exploiting different interaction methodologies are distributed with the complete software package.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Community Data Mining Approach for Surface Complexation Database Development

This paper presents a comprehensive data-to-model workflow, including a findable, accessible, interoperable, reusable (FAIR) community sorption database (newly developed LLNL Surface Complexation/Ion Exchange (L-SCIE) database) along with a data fitting workflow to efficiently optimize surface complexation reaction constants with multiple surface complexation model (SCM) constructs. This workflow serves as a universal framework to mine, compile, and analyze large numbers of published sorption data as well as to estimate reaction constants for parameterizing reactive transport models. Here the framework includes (1) data digitization from published papers, (2) data unification including unit conversions, and (3) data-model integration and reaction constant estimation using geochemical software PHREEQC coupled with the universal parameter estimation code PEST. We demonstrate our approach using an analysis of U(VI) sorption to quartz based on a first L-SCIE implementation, concluding that a multisite SCM construct with carbonate surface species yielded the best fit to community data. Surface complexation reaction constants extracted from this approach captured all available sorption data available in the literature and provided insight into previously published reaction constants and surface complexation model constructs. The L-SCIE sorption database presented herein allows for automating this approach across a wide range of metals and minerals and implementing novel machine learning approaches to reactive transport in the future.

38 RADIATION CHEMISTRY, RADIOCHEMISTRY, AND NUCLEA↗

DancePartner: Python Package to Mine Multiomics Relationship Networks from Literature and Databases

A goal of multi-omics experiments is to understand how mechanistic molecular biology is altered between conditions, typically a control group and experimental groups. Oftentimes this involves studying changes in biomolecule relationships (e.g. interactions, metabolic relationships) of several types of biomolecules (e.g. proteins, lipids, metabolites). Though several databases contain relationships between biomolecules, understudied species may have little to no relationship information in databases and thus must be mined from literature. There are several challenges to literature mining, including automated full-text extraction, duplicate biomolecule term collapsing, and implementing complex machine learning tools. To make relationship extraction more accessible to the community, a python package called DancePartner was developed to allow for the extraction of relationships from literature and databases, with functions to map biomolecule synonyms to standardized identifiers and visualize and characterize the resulting multi-omics network. Here, in this study, an example dataset involving Caenorhabditis elegans is presented, where relationships are mined from 1443 publications using DancePartner. These relationships are combined with relationships from KEGG, WikiPathways, UniProt, and LipidMaps, and visualized.

BERT↗

Global root traits (GRooT) database

Motivation: Trait data are fundamental to the quantitative description of plant form and function. Although root traits capture key dimensions related to plant responses to changing environmental conditions and effects on ecosystem processes, they have rarely been included in large-scale comparative studies and global models. For instance, root traits remain absent from nearly all studies that define the global spectrum of plant form and function. Thus, to overcome conceptual and methodological roadblocks preventing a widespread integration of root trait data into large-scale analyses we created the Global Root Trait (GRooT) Database. GRooT provides readyto- use data by combining the expertise of root ecologists with data mobilization and curation. Specifically, we (a) determined a set of core root traits relevant to the description of plant form and function based on an assessment by experts, (b) maximized species coverage through data standardization within and among traits, and (c) implemented data quality checks. Main types of variables contained: GRooT contains 114,222 trait records on 38 continuous root traits. Spatial location and grain: Global coverage with data from arid, continental, polar, temperate and tropical biomes. Data on root traits were derived from experimental studies and field studies. Time period and grain: Data were recorded between 1911 and 2019. Major taxa and level of measurement: GRooT includes root trait data for which taxonomic information is available. Trait records vary in their taxonomic resolution, with subspecies or varieties being the highest and genera the lowest taxonomic resolution available. It contains information for 184 subspecies or varieties, 6,214 species, 1,967 genera and 254 families. Owing to variation in data sources, trait records in the database include both individual observations and mean values. Software format: GRooT includes two csv files. A GitHub repository contains the csv files and a script in R to query the database.

59 BASIC BIOLOGICAL SCIENCES↗

Fine-Root Ecology Database (FRED): A Global Collection of Root Trait Data with Coincident Site, Vegetation, Edaphic, and Climatic Data, Version 4.

To address the need for a centralized root trait database, we compiled the Fine-Root Ecology Database (FRED) from published and unpublished data sources. We have continued to add to the FRED database since the release of FRED 1.0 in 2017, followed by 2.0 in 2018, and 3.0 in 2021. This new release of FRED 4.0 now has 213,941 observations of 238 root traits, for a combined total of roughly 3.4 million data fields for root traits and ancillary data together. FRED 4.0 has 39.8% more root trait observations than FRED 3.0 and a 34.4% increase in unique data sources. This release of FRED 4.0 also includes significant increases in geographic regions that have long been underrepresented in global datasets, notably in the tropical low latitudes. Ancillary data on associated site, vegetation, edaphic, and climatic conditions from across the globe have also increased concurrently with root trait observations. FRED is focused on fine roots (traditionally defined as roots less than 2 mm in diameter), as coarse roots are studied using different methodology, often at very different scales, and have different traits and trait interpretations. Despite this fine-root focus, FRED accepts data collected from roots of all sizes and contains observations of many root classes including coarse roots. Data collection will continue for the foreseeable future. The FRED4_Entire_Database_2026.csv file is the flat csv data file for FRED 4.0, and the FRED4_dd.csv file is the data dictionary of all columns available in FRED, including column IDs, column names, definitions, and unit (where applicable).

54 ENVIRONMENTAL SCIENCES↗

NEWTS Argonne Geothermal Geochemical Database with CoDART

A database of geochemical compositions of aqueous species in potential geothermal resources. The NETL NEWTS team has formatted the original Argonne V2 database for Charge Balance, Input into OLI Studio, and Input into GWB. (Geochemist WorkBench) In addition, some missing species in the original database were predicted using machine learning techniques within CoDart software, a public ML software developed by the Nation Energy Technology Laboratory. We have made the Input into CoDart and one example output from CoDart available in this dataset.

Aqueous Chemistry↗

Classification of Cloud Particle Imagery and Thermodynamics (COCPIT): A New Databasing Tool for the Characterization of Cloud Particle Images Captured During DOE Field Campaigns

The Department of Energy for decades has explored the earth system and atmosphere through research and deployment of in-situ and remote sensing platforms during field campaigns. Among these datasets exists a vast supply of cloud particle images that provide visual insight into the complex microphysics in the clouds that span our globe. The millions of images collected over decades of deployments provides a unique opportunity to further our understanding of our atmosphere down to the crystal size. This work over the past 5 years has sought to organize these images into digestible datasets that can then be used by scientists to further our understanding of microphysics. A machine learning model was developed that categorizes over 1.5 million images across 11 weather events with over 90% accuracy according to particle type. The database was then extended to include dimensional characteristics of the particle as well as co-location of environmental properties, such as temperature and water content. Then, to initialize the connection between these data and our understanding of how crystals form and grow, weather research and forecasting simulations were run to generate the growth histories of the classified crystals. This research culminates with 2 databases per event: (1) a database of all classified crystals and their dimensional and environmental properties and (2) simulated growth histories of each crystal. Finally, a user interface was created to allow researchers to explore data statistics.

54 ENVIRONMENTAL SCIENCES↗

Fine-Root Ecology Database (FRED): A Global Collection of Root Trait Data with Coincident Site, Vegetation, Edaphic, and Climatic Data, Version 3.

To address the need for a centralized root trait database, we compiled the Fine-Root Ecology Database (FRED) from published and unpublished data sources. We have continued to add to the FRED database since the release of FRED 2.0 in 2018, and a new version of FRED is now available. FRED 3.0 has more than 150,000 observations of more than 330 root traits, with data collected from more than 1400 data sources. FRED 3.0 has 45% more root trait observations than FRED 2.0, particularly in the categories of root anatomy, morphology, and microbial associations; ancillary data on associated site, vegetation, edaphic, and climatic conditions from across the globe have also increased concurrently. FRED is focused on fine roots (traditionally defined as roots less than 2 mm in diameter), as coarse roots are studied using different methodology, often at very different scales, and have different traits and trait interpretations. However, FRED accepts data collected from roots of all sizes, and already contains several observations of coarse roots. Data collection will continue for the foreseeable future.

54 ENVIRONMENTAL SCIENCES↗

A market-oriented database design for critical material research

Material databases are important tools to provide and store information from material research. Rising concerns about supply-chain risks to raw materials presents a need to incorporate raw-material market and end-use application data, beyond basic chemical and physical properties, into a material database. One key challenge for researchers working on critical materials is information scarcity and inconsistency. This paper introduces, as a result of a two-year project, a critical-material commodity database (CMCD) incorporated with a low-code web-based platform that allows easy access for users and simple updates for the authors. The main goal of this project was to educate material scientists on the applications having the most impact on the supply chain and current industrial specifications/markets for each application. The objective was to provide material researchers with harmonized information so that they could gain a better understanding of the market, focus their technologies on an application with a high potential for commercialization, and better contribute to supply-chain risk reduction. While the goal was met with high receptivity, several limitations stemmed from query design, distribution platform, and quality of data source. To overcome some of these limitations and expand on CMCD's potential, we are building a public webpage with an improved interface, better data organization, and higher extensibility.

36 MATERIALS SCIENCE↗

Efficient hemodynamic event detection utilizing relational databases and wavelet analysis

Development of a temporal query framework for time-oriented medical databases has hitherto been a challenging problem. We describe a novel method for the detection of hemodynamic events in multiparameter trends utilizing wavelet coefficients in a MySQL relational database. Storage of the wavelet coefficients allowed for a compact representation of the trends, and provided robust descriptors for the dynamics of the parameter time series. A data model was developed to allow for simplified queries along several dimensions and time scales. Of particular importance, the data model and wavelet framework allowed for queries to be processed with minimal table-join operations. A web-based search engine was developed to allow for user-defined queries. Typical queries required between 0.01 and 0.02 seconds, with at least two orders of magnitude improvement in speed over conventional queries. This powerful and innovative structure will facilitate research on large-scale time-oriented medical databases.

NASA Discipline Cardiopulmonary↗