Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “relational database”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

The Baghdad Atlas: A relational database of inelastic neutron-scattering (n,n ' γ) data

A relational database has been developed based on the original (n,n'γ) work carried out by A. M. Demidov et al., at the Nuclear Research Institute in Baghdad, Iraq (Demidov et al., 1978) for 105 independent measurements comprising 76 elemental samples of natural composition and 29 isotopically-enriched samples. The information from this Atlas includes: γ-ray energies and relative intensities; nuclide and level data corresponding to the residual nucleus and meta data associated with the target sample that allows for the extraction of the flux-weighted (n,n'γ) cross sections for a given transition relative to a defined value. The optimized angular-distribution-corrected fast-neutron flux-weighted partial γ-ray cross section for the production of the 846.8-keV 21+→0gs+γ-ray transition in 56Fe, determined to be $\langle$σγ$\rangle$=143(29) mb, is used for this purpose. However, different values for the adopted cross section can be readily implemented to accommodate user preference based on revised determinations of this quantity. The Atlas (n,n'γ) data has been compiled into a series of CSV-style ASCII data sets and a suite of Python scripts have been developed to build and install the database locally. The database can then be accessed directly through the SQLite engine, or using alternative methods such as the Jupyter Notebook Python-browser interface. Several examples exploiting different interaction methodologies are distributed with the complete software package.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Symmetry relation database and its application to ferroelectric materials discovery

To investigate the displacive phase transition at the atomic scale, we have implemented a numerical algorithm to automate the detection of the symmetry relations between any two candidate crystal structures. Using this algorithm, here we systematically screen all possible polar–nonpolar structure pairs from the Materials Project database and establish a library of ~4500 pairs that can be connected through a continuous phase transition with small atomic displacements. From this database, we identify several new ferroelectric materials. In addition, the database may also be used in other areas, such as material structure prediction and new materials discovery.

36 MATERIALS SCIENCE↗

SoK: What does it Mean to Benchmark Database Forensics?

Relational Database Management Systems are the backbone of modern enterprises and public-sector services, and are thus frequent targets of security incidents, insider threats, and thorough regulatory audits. Consequently, databases have become key sources of digital evidence, requiring investigators to reconstruct past activity from audit logs, transaction logs, and backups. Although benchmarking frameworks such as those developed by the Transaction Processing Performance Council (TPC) are widely used to evaluate database performance, they do not capture forensic requirements such as evidentiary completeness, tamper-evidence, chain of custody, or regulatory compliance under GDPR and CCPA. This survey examines the emerging domain of forensic database benchmarking. We gathered prior research on database forensics, secure logging, and tamper-evident data structures; we analyze modern forensic-ready features in commercial and open-source systems (SQL Server Ledger, Oracle Blockchain Tables, PostgreSQL pgAudit, Db2 Audit, Aurora Database Activity Streams, Oracle Real Application Security and IBM Guardium) and assess why existing benchmarks are insufficient. We propose forensic workloads, metrics, and methodologies that incorporate adversarial stressors, deleted-record recovery, and backup analysis. We also identify open research problems and call for a community-driven forensic benchmark suite. The result is an idea for evaluating not only database performance but also forensic soundness, bridging the gap between system engineering, compliance, and digital investigations.

Lenard, Ben↗

Generic National Nuclear Forensics Library Implementation

This document provides instructions for the database implementation for national nuclear forensic library (NNFL) using a relational database. While there are many software options that can be used to implement an NNFL database, in this document we reference an Oracle database for the implementation. The principles related here can be applied to any relational database structure. The design revolves around two main types of data: Samples and Results. Samples are materials or descriptions of materials; results are the results of analyses of those samples. All additional tables exist as a result of the normalization process. The remaining tables/entities have keys that are referenced by the Result/Sample tables via foreign key constraints. This allows for the enforcement of different relationships between the tables such as one-to-one, one-to-many, many-to-many. This process helps to ensure that valid data is entered, optimizes database performance and avoids data redundancies.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

TenSQL v.2023.01.20

SAND2024-02520O Tensor SQL Database performs certain relational database management system/structured query language (RDBMS/SQL) queries faster than is possible using existing state-of-the-art databases. Tensor SQL is optimized for sparse linear algebra problems and other whole-table queries. Currently, the database is being used only for research and program development. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Roose, Jonathan↗

CS-Studio Alarm System Based on Kafka

The CS-Studio alarm system was originally based on a relational database and the Apache ActiveMQ message service. The former was necessary to store configuration and state, while the latter communicated state updates and user actions. In a recent update, the combination of relational database and ActiveMQ have been replaced by Apache Kafka. We present how this simplified the implementation while at the same time improving performance.

Kasemir, Kay↗

EXFOR-NSR PDF database: a system for nuclear knowledge preservation and data curation

Current needs of nuclear science and technology include complete, well-documented, and easily verifiable nuclear data. The complete data records require supporting nuclear bibliography, presently stored in dedicated libraries, in addition, to actual data. Additionally, experimental nuclear reaction data (EXFOR) and Nuclear Science References (NSR) databases contain compilations based on primary (journals) and secondary (conference proceedings, theses, preprints, etc.) publications, and data received from authors via private communications. The secondary library materials and private communications often represent a bottleneck for nuclear data verification, compilation, evaluation, and dissemination activities. To address this issue, bibliographic materials were scanned into PDF (Portable Document Format) files and uploaded in a relational database. The traditional scope of nuclear databases that includes meta-data and numbers derived from data in specialized formats was broadened to accommodate the large volumes of original nuclear data publications. The complete PDF publication files were stored in a relational database as Binary Large OBjects (BLOB). This unique collection of nuclear data compilations and supporting publications generate many opportunities for machine learning applications. The Web interfaces for authorized and public access to the EXFOR-NSR nuclear publications database were implemented at the U.S. National Nuclear Data Center, https://www.nndc.bnl.gov/ and IAEA Nuclear Data Section, https://www-nds.iaea.org/ . The current system is complementary to major nuclear libraries and narrowly focused on nuclear data compilation and evaluation procedures. The contents of the PDF database, details of implementation, and Web interface are described. New capabilities for data curation, knowledge preservation, worldwide dissemination, and natural language processing (NLP) applications are given.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

An uncertainty-focused database approach to extract spatiotemporal trends from qualitative and discontinuous lake-status histories

Changes in lake status are often interpreted as palaeoclimate indicators due to their dependence on precipitation and evaporation. The Global Lake Status Database (GLSDB) has since long provided a standardised synopsis of qualitative lake status over the last 30,000 14C years. Potential sources of uncertainty however are not recorded in the GLSDB. Here we present an updated and improved relational-database framework that incorporates uncertainty in both chronology and the interpretation of palaeoenvironmental data. The database uses peer-reviewed palaeolimnological studies to produce a consensus on qualitative lake-status histories, whose chronologies are revised and standardized through the recalibration of radiocarbon dates and the application of Bayesian age-depth modelling for stratigraphic archives. Quantitative information on absolute water-level elevation is preserved if available from geomorphological sources. We also propose a new probabilistic analytical framework that accounts for these uncertainties to reconstruct synoptic, integrated environmental signals. The process is based on a Monte Carlo algorithm that iteratively samples individual lake-status histories within the limits of their uncertainties to produce many possible scenarios. We then use Recursively-Subtracted Empirical Orthogonal Function analysis to extract dominant patterns of lake-status variability from these scenarios. As a proof of concept, we apply this framework to 67 sites in eastern and southern Africa whose lake-status histories cover part of the late Pleistocene and/or Holocene. We show that, despite the sometimes large temporal and interpretation uncertainties, and the inclusion of highly discontinuous lake-status time series, identifying the major known millennial-scale climatic phases during the last 20,000 years is possible. Our framework was also able to identify an antiphased response between the lake basins in eastern and interior southern Africa to these changes. Here, we propose that our new database and methodology framework serves as a template for efficient lake-status data synthesis, encourages the incorporation of lake-status data in palaeoclimate syntheses, and expands the possibilities for the use of such data in the evaluation of climate models.

58 GEOSCIENCES↗

TEACHING AN OLD ACCELERATOR NEW TRICKS

The Argonne Tandem Linac Accelerator System (ATLAS) has been a National User Facility since 1985. In that time, many of the systems that help operators retrieve, modify, and store beamline parameters have not kept pace with the advancement of technology. Development of a new method of storing and retrieving beamline parameters resulted in the testing and installation of a time-series database as a potential replacement for the traditional relational database. InfluxDB was selected due to its self-hosted Open-Source version availability as well as the simplicity of installation and setup. A program was written to periodically gather all accelerator parameters in the control system and store them in the time-series database. This resulted in over 13,000 distinct data points, captured at 5-minute intervals. A second test captured 35 channels on a 1-minute cadence. Graphing of the captured data is being done on Grafana, an Open-Source version is available that co-exists well with InfluxDB as the back-end. Grafana made visualizing the data simple and flexible. The testing has allowed for the use of modern graphing tools to generate new insights into operating the accelerator, as well as opened the door to building large data sets suitable for Artificial Intelligence and Machine Learning applications.

Novak, D.↗

Specifications of Legacy U(Pu)Zr Metallography Data

The DOE Nuclear Energy Advanced Reactor Technologies (ART) Program has supported the creation of several databases with information describing the safety performance of fast reactors, components, and fuels. This growing collection of legacy experimental data, operating data, and analysis is available online to registered users. Metallography data is one of the most important types of PIE data being collected, organized and stored in several ART Fast Reactor Databases (https://frdb.ne.anl.gov), including the Fuels Irradiation & Physics Database (FIPD), Out-of-Pile Transient Database (OPTD), and TREAT (the Transient Reactor Test Facility) Experimental Relational Database (TREXR). These databases contain three main sets of metallography data. The first set is the metallography data from Experimental Breeder Reactor-II (EBR-II) and Fast Flux Test Facility (FFTF) irradiated fuel pins measured in the Hot Fuel Examination Facility (HFEF); the second set is the metallography data from EBR-II irradiated fuel pins measured in the Alpha-Gamma Hot Cell Facility (AGHCF); the third set is the metallography data from transient tested fuel pins (the transients tests include the out-of-pile tests and TREAT tests) measured in AGHCF. Since both the second and third sets of data were measured in AGHCF, they are governed by the same specification. The metallography data in the databases are digital images scanned from either positive or negative photos. The quality of the images, including their resolution and contrast, relies on the preserved quality of the pictures and the scanning conditions. To analyze the microstructure of a fuel pin, a series of preparatory steps must be undertaken. These include sectioning, epoxy mounting, mechanical grinding and polishing, and often etching. The resulting samples were then transferred to a secondary hot cell, or glovebox (depending the strength of radiation field) for microscopy examination. The specifications provided herein focus on the metallography examinations. The sample grinding, polishing and etching processes are also discussed. The procedures involving sectioning and epoxy mounting are out of the scope of the current specification; details can be found in the corresponding operation manuals. If more data are collected and added to the ART Fast Reactor Databases, this specification will be updated to accommodate the additional data.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

HIV Molecular Immunology 2025

HIV Molecular Immunology is a companion volume to HIV Sequence Compendium. This publication, the 2025 edition, is the PDF version of Los Alamos Na tional Laboratory’s web-based HIV Molecular Immunology Database (https://www.hiv.lanl.gov/content/ immunology/). The web interface for this relational database has many search interfaces for HIV immunological in formation, as well as interactive tools to help immunologists design reagents and interpret their results.

59 BASIC BIOLOGICAL SCIENCES↗

Determination of Scale Bar for AGHCF Metallography Data

The DOE Nuclear Energy Advanced Reactor Technologies (ART) Fast Reactor Program (FRP) has supported the development of several databases containing information on the safety performance of fast reactors, components, and fuels. This growing collection of legacy experimental data, operating data, and analysis is available online to registered users. Metallography data represents one of the most critical types of post-irradiation examination (PIE) data being collected, organized, and archived in several ART Fast Reactor Databases (https://frdb.ne.anl.gov), including the Metallic Fuels Irradiation & Physics Database (FIPD), Out-of-Pile Transient Database (OPTD), and TREAT (the Transient Reactor Test Facility) Experimental Relational Database (TREXR). These databases contain three principal sets of metallography data. The first set comprises metallography data from Experimental Breeder Reactor-II (EBR-II) and Fast Flux Test Facility (FFTF) irradiated fuel pins examined in the Hot Fuel Examination Facility (HFEF). The second set consists of metallography data from EBR-II irradiated fuel pins examined in the Alpha-Gamma Hot Cell Facility (AGHCF). The third set includes metallography data from transient-tested fuel pins (including both out-of-pile furnace tests and TREAT tests) examined in AGHCF. Since both the second and third sets were generated in AGHCF, they are governed by identical specifications. The metallography data in the databases consist of digital images scanned from either positive or negative photographic films. To analyze the microstructure of a fuel pin, a series of preparatory steps are required, including sectioning, epoxy mounting, mechanical grinding and polishing, and etching. Following sample preparation, specimens are transferred for metallographic examination. The AGHCF and HFEF metallography data were generated using optical microscopes manufactured by Leitz and Bausch and Lomb (B&L). Images were recorded on Polaroid film at preset magnifications. Magnification verification for the Leitz and B&L metallographs was conducted every two months prior to 1989 and at least every six months from 1989 through the conclusion of the IFR program. Magnifications determined from imaging of microslide standards were compared to the instrument settings for magnifications ranging from 50× to 500×. If the magnifications determined from standards deviated from the instrument settings, adjustments were made to the bellows extension until agreement was achieved. The specifications for AGHCF and HFEF legacy metallography data have been established based on available hard-copy and digital records, most of which have been incorporated into the data repositories associated with FIPD, OPTD, and TREXR. Detailed specifications including hard-copy records, digitized records, cutting diagrams and sectioning schemes, high-magnification photographs, photomosaics (composites), information tags, scale bars, and magnification verification procedures can be found in a separate report.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

The Materials Provenance Store

Abstract We present a database resulting from high throughput experimentation, primarily on metal oxide solid state materials. The central relational database, the Materials Provenance Store (MPS), manages the metadata and experimental provenance from acquisition of raw materials, through synthesis, to a broad range of materials characterization techniques. Given the primary research goal of materials discovery of solar fuels materials, many of the characterization experiments involve electrochemistry, along with optical, structural, and compositional characterizations. The MPS is populated with all information required for executing common data queries, which typically do not involve direct query of raw data. The result is a database file that can be distributed to users so that they can independently execute queries and subsequently download the data of interest. We propose this strategy as an approach to manage the highly heterogeneous and distributed data that arises from materials science experiments, as demonstrated by the management of over 30 million experiments run on over 12 million samples in the present MPS release.

36 MATERIALS SCIENCE↗

Insert Modeling in UNF ST&DARDS

The Used Nuclear Fuel-Storage, Transportation & Disposal Analysis Resource and Data System (UNF-ST&DARDS) is a software tool that integrates a used nuclear fuel (UNF) or spent nuclear fuel (SNF) relational database and key analysis capabilities to simplify and automate numerous UNF management and fuel cycle–related activities. UNF-ST&DARDS is being developed for the US Department of Energy’s Office of Nuclear Energy Spent Fuel and Waste Disposition program. UNF-ST&DARDS provides an integrated framework that uses advanced modeling and simulation to predict the behavior of SNF over the timescales associated with permanent disposal in a geologic repository. After leaving the spent fuel pool, SNF is transferred to dry storage in a dual-purpose canister (DPC). DPCs are considered “dual purpose” because they are designed for both storage and transportation, removing the need to transfer the fuel to a separate transportation cask. However, much research has been conducted investigating the feasibility of directly disposing of DPCs in geologic repositories. Direct disposal of DPCs could reduce worker exposure during repackaging, reducing the amount of low-level waste from the discarded DPCs and potentially saving billions of dollars. Therefore, direct disposal of as-loaded DPCs is desirable if it can be done safely.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

SQL and NoSQL Databases for Cyber Physical Production Systems in Internet of Things for Manufacturing (IoTfM)

Abstract In this paper, the design and performance differences between Relational Database Management Systems (RDBMS) and NoSQL Database Systems are examined, with attention to their applicability for real-world Internet of Things for manufacturing (IoTfM) data. While previous work has extensively compared SQL and NoSQL for both generalized and IoT uses, this work specifically examines the tradeoffs and performance differences for manufacturing applications by using a high-fidelity data set collected from a large US manufacturing firm. Growing an IoT system beyond the pilot stage requires scalable data storage; this work seeks to determine the impact of selected database systems on data write performance at scale. Payload size and message frequency were used as the primary characteristics to maintain model fidelity in simulated clients. As the number of simulated asset clients grow, the data write latency was calculated to determine how both database systems’ performance were affected. To isolate the RDBMS and NoSQL differences, a cloud environment was created using Amazon Web Services (AWS) with two identical data ingestion pipelines: writing data to an RDMBS (1) using AWS Aurora MySQL, and (2) using AWS DynamoDB NoSQL. The findings may provide guidance for further experimentation in large-scale manufacturing IoT implementations.

Gamero, David↗

Trends in Light and Temperature Sensitivity Recommendations among Licensed Biotechnology Drug Products

Inherent structural and functional properties of biotechnology-derived therapeutic biologics make them susceptible to light- and temperature-induced degradation and consequently can influence their quality. Photosensitivity of therapeutic proteins continues to be examined, but the commonalities and trends of storage conditions and information about light and temperature sensitivity among currently licensed therapeutic proteins has not been previously surveyed. Using a comprehensive and relational database approach, we conducted a scientific survey of all licensed biotechnology-derived drug products with the goal of providing evidence-based information about recommended storage conditions of formulations sorted by light- and temperature-related attributes as described for each product at licensure. We report the prevalence of indications for light and temperature sensitivity in formulations categorized by their presentation type, number of doses, container type, dosage form and active molecule type. We also report the storage temperature range across formulations and diluents for reconstitution and dilution. Formulations with excipients that potentially facilitate light-induced and thermal degradation were also noted. The result of our analysis indicates that light and temperature sensitivity are prevalent across therapeutic protein formulations. However, when a formulation is reconstituted or diluted, both light and temperature sensitivity are less clear. In addition, light and temperature sensitivity are more well defined in liquid formulations than lyophilized powder formulations, and more well defined in products manufactured in autoinjectors, prefilled-syringes, and pens than products in vials. Overall, our report provides a data-driven summary of storage conditions among therapeutic protein formulations to support the development of future biologic drug products.

60 APPLIED LIFE SCIENCES↗

Towards Auto-Generated Data Systems

After decades of progress, database management systems (DBMSs) are now the backbones of many data applications that we interact with on a daily basis. Yet, with the emergence of new data types and hardware, building and optimizing new data systems remain as difficult as the heyday of relational databases. In this paper, we summarize our work towards automating the building and optimization of data systems. Drawing from our own experience, we further argue that any automation technique must address three aspects: user specification, code generation, and result validation. We conclude by discussing a case study using videos data processing, along with opportunities for future research towards designing data systems that are automatically generated.

Computer Science↗

OPTOM: Optimization of Parabolic Trough - Operations & Maintenance

The US Department of Energy’s SunShot goals look to reduce the cost of Concentrating Solar Power (CSP) technology to 5¢/kWh for baseload plants. This is about a 50% reduction from current costs. To achieve this cost target, a significant reduction in operation and maintenance (O&M) costs of 40 to 50% is likely needed. Advances are needed in the O&M practices of CSP plants if the technology is to achieve the SunShot cost goals. Digitization of plant performance and O&M data has become a new best practice in the world of renewable energy asset management. Owners and operators of large photovoltaic and wind power plants are working to digitize performance and O&M data at their existing assets, to improve their management of the facilities, to increase performance, reduce O&M costs, and lower the overall life cycle cost of ownership. CSP power plants are behind the curve of other technologies on the digitization of plant information to aid in the plant asset management. This project directly addresses the objective of digitizing the O&M data of the solar field, focusing on three areas: 1) creating a framework for sharing data and information, 2) creating a system for monitoring and managing the maintenance of the solar field collectors, and 3) developing analytic tools to identify issues in the solar field. According to the NREL CSP Best Practices Study, the current practice at many CSP plants is to rely on paper lists, spreadsheets, and email for monitoring and managing problems and maintenance in the solar field. The key element to digitize solar field O&M is the creation of a centralized data archive that all users and systems can interface with. This project developed a centralized relational database framework that allows users and applications to access and share data. Conventional power plants utilize Computerized Maintenance Management Systems (CMMSs) to track the corrective, preventive (scheduled), and predictive maintenance of equipment and subsystems in the power plant. CSP plants use these systems in the power block, but while these systems specialize at tracking maintenance on up to thousands of pieces of equipment, they are not well suited for tracking the tens or hundreds of thousands of components in large commercial CSP or photovoltaic solar fields. In this project we developed a new software application referred to as FieldStatus (TM). This is a specialized database program that is used to track the status of each collector and its components. This application is designed to complement the existing CMMS to enable improved tracking and management of maintenance activities in the solar field. One of the major maintenance tasks for solar fields is maintaining the cleanliness of the mirrors. Although seemingly a relatively straight forward task, it has often proven challenging to maintain high levels of cleanliness in an efficient and cost-effective manner. This project developed new tools and metrics for monitoring and optimizing solar field cleaning resources and overall solar field cleanliness.

14 SOLAR ENERGY↗