Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “database management”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

Used Nuclear Fuel Management Using the Next Generation System Analysis Model

The U.S. Department of Energy (DOE) is leading the National effort to manage the back end of the nuclear fuel cycle, encompassing the safe transportation, storage/staging, and/or eventual disposal of used nuclear fuel (UNF) and high-level radioactive waste. The Next Generation System Analysis Model (NGSAM) is DOE’s discrete-event, agent-based simulation tool designed to model the full life cycle of UNF from reactor discharge to final disposal. NGSAM supports the DOE Office of Spent Fuel and High-Level Waste Disposition by enabling a detailed, scenario-based analysis of logistics, infrastructure, and shipping strategies. NGSAM replaces legacy models with a modern, flexible platform built on Repast Simphony and enhanced by the Process Analysis Tool. NGSAM simulates the movement and interaction of individual fuel assemblies with system components such as canisters, casks, railcars, and facilities. The model integrates with the Java Transportation Operations Model to plan and execute transportation scenarios, supporting both constrained and unconstrained resource allocation. Key features include customizable allocation and acceptance algorithms, detailed facility-level operations, and a Quick Edit tool for rapid scenario adjustments. NGSAM supports multimodal transportation modeling (e.g. rail, road, barge) and provides comprehensive cost, schedule, and infrastructure data. NGSAM utilizes data from sources such as DOE’s STANDARDS UNF database and DOE’s Stakeholder Tool for Assessing Radioactive Transportation, while also allowing user-defined inputs for scenario customization. NGSAM enables stakeholders to evaluate complex UNF management strategies, assess system performance under varying assumptions, and inform decision making for future infrastructure investments. Its modular architecture and integration with other Integrated Waste Management System tools make it a critical asset for planning the safe and efficient disposition of the Nation’s growing UNF inventory.

Craig, Brian [Argonne National Laboratory (ANL)]↗

M3SF-24LL010302062-NEA-TDB Management and International Collaborations in Sorption and Thermodynamic Modeling

This progress report (Level 3 Milestone Number M3SF-24LL010302062) summarizes research conducted at Lawrence Livermore National Laboratory (LLNL) within the Crystalline International Collaborations Work Package Number SF-24LL01030206. The activity is focused on our long-term commitment to engaging our partners in international nuclear waste repository research. This includes participation in the Nuclear Energy Agency Thermochemical Database (NEA-TDB) Project and development of methodologies for integrating US and international thermodynamic databases for use in SFWST Generic Disposal System Assessment (GDSA) efforts. A continuing focus for FY24 efforts is to support the US participation in the NEA-TDB effort. The focus of FY24 activities was the development of an agreement for a Phase 7 activity that will start in Q1 of 2025. Mavrik Zavarin is now the US representative on both the Management Board and the Executive group to the NEA-TDB. He is also the POC for the Cements State of the Art Report that is undergoing peer review in FY24. In FY24, we used our position on the NEA-TDB MB and EG to facilitate the integration of NEA-TDB thermochemical data with LLNL’s SUPCRTNE thermodynamic database that supports the SFWST GDSA activities. This effort is coordinated with the Argillite work package SUPCRTNE database development efforts (Wolery, 2024). The goal is to provide a downloadable database that will be hosted on LLNL’s thermodynamics website which incorporates NEA-TDB data into the LLNL database where appropriate. We also began engagement with the EURAD2 program that was initiated in FY24 by our European collaborators at the Karlsruhe Institute of Technology (KIT), Germany. The primary focus of the engagement is with WP20: DITUSC Thermodynamic database evaluation program. A kickoff meeting for this activity is planned for early FY25. Finally, we have been selected to co-host (with Clemson University) the International Conference on Chemistry and Migration Behaviour of Actinides and Fission Products in the Geosphere in 2025 (Migration2025). The meeting will be held September 21-26, 2025, in New Orleans, Louisiana, and will focus on international efforts to understand the risks of radionuclide releases into the environment. This central focus of this conference is on international efforts to develop safe disposal options for nuclear wastes. As such, we are developing a theme focused on US underground nuclear waste repository science.

58 GEOSCIENCES↗

Charpy Impact Characterization of Surveillance Specimens Harvested from Palisades High Fluence A-60 Capsule

Located on the shores of Lake Michigan, the Palisades Nuclear Generating Station (PNGS) was a nuclear power plant that operated in Covert Township, Michigan. The plant had a single pressurized water reactor that produced electricity for the region. The PNGS was shut down in 2022 after more than four decades of service. The PNGS included in its surveillance program a surveillance capsule, designated A-60, containing specimens of a weld metal with nickel content of about 1.36 wt% and copper content of about 0.25 wt%. The capsule was removed from its surveillance position in early 1995 and has been resident in the spent fuel pool since that time. This capsule was irradiated to a fluence of 1.96×10 20 n/cm 2 (E>1MeV) that is equivalent for more than 150 effective full power years (EFPYs) for the US reactor pressure vessel (RPV) fleet. The material is also of special interest because of its very high nickel content and potential for development of NiMnSi (nickel-manganese-silicon) precipitates. Combination of very high fluence and very high Ni and Cu content makes the material in this capsule of great interest as benchmark for currently developing embrittlement trend curves (ETC) aiming to predict embrittlement at high fluences. Multi-year efforts by the Light Water Reactor Sustainability (LWRS) Program personnel from the Materials Research Pathway (MRP) to harvest this capsule finally succeeded in 2023. As a result, the Westinghouse Electric Company (WEC) through contract with Oak Ridge National Laboratory (ORNL) came to PNGS site, retrieved the A-60 capsule, brought it to WEC Churchill hot cell facility, opened the capsule and sent all surveillance specimens in the capsule to ORNL for future characterization by July 2023. Total of 9 tensile and 48 Charpy specimens were inventoried in the ORNL hot cells. Testing plan has been developed based on available specimens. It includes hardness, tensile, Charpy impact, Mini-CT fracture toughness testing, in-situ thermal annealing, and microstructural characterization, including Atom Probe Tomography (APT). Charpy impact testing has been completed and results are presented in this report. Moreover, negotiations with Pressurized Water Reactors Owners Group (PWROG) and Westinghouse resulted in Westinghouse donating to ORNL a piece of archive weld and base metal such that unirradiated characterization of these materials can be performed as part of this project. Comparison of measured hardening and embrittlement has been performed and data are compared to available large database.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

Datashare

Datashare facilitates communication and data sharing within local networks in potentially dangerous situations such as an explosive ordnance disposal. During such events, there is a need to transmit information rapidly around the incident area. It is a distributed database that does not require an internet connection for operation. In addition, Datashare interfaces with XTK and other software applications, allowing for seamless integration and data management. Datashare supports video calls over the network, enabling real-time communication among users. This software serves to organize, package, and share between responders on location and export data to those off location. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Eldridge, Bryce [Sandia National Lab. (SNL-CA), Li↗

Assessment of Bird Strike Likelihood to Refine Bird Strike Risk Models

In its most basic form, bird strike risk is comprised of a frequency component that reflects the likelihood of a collision and a severity component that reflects the cost (monetary or otherwise) of the incident. The bird strike risk model currently used by United State Department of Agriculture (USDA) Wildlife Services to evaluate the risk posed by individual bird species at airports and establish priorities for management was developed in 2018. The model uses airport-specific data on the number of reported strikes for a species recorded in the Federal Aviation Administration (FAA)’s National Wildlife Strike Database as a measure of frequency and the species’ relative hazard score as a measure of severity. The model was tested against independent data, found to perform well overall, and is being implemented widely across the United States. However, the model has limitations, including that species known to pose risk to aircraft locally, but not present in the strike record database, are not reflected as a major component of risk. Standard bird survey methodology commonly used at airports (e.g. point counts or transects) potentially can be used to complement wildlife strike records to calculate frequency or relative abundance of species. However, these methods generally focus on airport-wide population estimation and often ignore vital information that contributes to the true likelihood of a strike, such as use of runway protection zones and other critical areas, and spatial and temporal overlap with departing or approaching aircraft. As such, a more detailed understanding of space use by birds across landcovers and population fluctuations across the year is needed to accurately estimate the likelihood of bird strikes at airports. In this manuscript, we will review the extant risk model, including a discussion on its limitations. We then discuss approaches for refining our understanding of strike likelihood and briefly touch on needs for estimating probability of strike severity (cost).

bird strike, aircraft collision, damage by wildlif↗

A customizable data management framework for high-repetition-rate high-energy-density science

The high-energy-density (HED) physics community is moving toward a new paradigm of high-repetition-rate (HRR) operation. To fully leverage the scientific power of HRR HED facilities, all of the components of each subsystem (laser, targetry, and performance diagnostics) must be connected and synchronized in a reliable and robust manner while the data acquired are tagged and archived in real time. To this end, GA has begun developing a generalized NoSQL-database framework, the MongoDB repository for information and archiving. An organizational strategy has been developed that shifts HED data organization from a shot-based to a diagnostic-based approach in order to increase archival and retrieval efficiency that lends itself to optimization applications. This work is a first step in pushing HRR HED science toward data management solutions that emphasize machine actionability and aim to stimulate community engagement to define data standards in HED science.

Instruments & Instrumentation↗

An Open-source Llm Enhanced-tool Specialized In Helping Moose Related Problems And Tasks

MOOSEenger is an open-source, terminal-first chat application for the MOOSE ecosystem that couples specialized parsing of MOOSE documentation and “.i” input files with retrieval-augmented generation to deliver grounded answers about multiphysics modeling and workflows. It includes dedicated readers for MOOSE-style HTML and a pyhit-based parser that uses the MOOSE syntax tree to preserve block structure and attach retrieval metadata. A data-ingestion pipeline performs semantic chunking into atomic facts and stores them hierarchically in a local Chroma vector database that maintains parent–child relationships across documents; the system can ingest directories, individual files, and single-page web content, and it provides CRUD operations (insert, update, delete) to manage the corpus. At query time, relevant chunks are embedded, retrieved, and fused into the model context, with interactive features such as token streaming, persistent chat history, and dynamic RAG (retrieval triggered by user input or intermediate model output). Deployment is flexible: MOOSEenger runs with local Ollama models or remote Hugging Face/OpenAI backends—typically coordinating generation, lightweight tagging/summarization, and embeddings across three models—and it also supports a server mode and integration with the VS Code Continue interface.

Li, Mengnan [Idaho National Laboratory (INL), Idah↗

EVs@Scale Next-Gen Profiles - EV Profile Capture 2024

As part of the U.S. DOE EVs@Scale consortium Next-Gen Profiles (NGP) project, the profile capture and analysis of production electric vehicles undergoing high power charging (HPC) is conducted over a wide range of conditions to explore variance and performance. Charge session parameters are collected from both the electric vehicle (EV) and electric vehicle supply equipment (EVSE) at a rate of 10Hz and entered into a time-series database for analysis. These charge profiles are captured under nominal and off-nominal conditions, exploring the impact of battery state of charge (SOC), battery temperature, vehicle condition, smart charge management (SCM), and EVSE limitations. Nominal conditions are defined to be ideal conditions that should transfer the maximum allowable energy in the minimum possible amount of time. Nominal condition profiles are compared across EVs to characterize state-of-the-art EV charging performance against one another. Off-nominal condition profiles are compared against its nominal condition profile counterpart to highlight the variance across less desirable starting conditions within a single EV. This EV Profile Capture 2024 report stands as an update from the EV Profile Capture 2023 report to include the additional EV & EVSE assets tested and analyzed in 2024. The major updates within this report include the addition of three next-generation electric vehicles, added test cases, and further analysis. This expansion of analysis includes power profiles, power distribution, quantifying SOC, energy and range performance, EVSE limitation impacts, boost converter performance, etc. Additionally, NGP time-series data has been used as input towards three national laboratory-led grid modelling efforts: ANL’s IEEE-37 HIL model, INL’s Caldera model, and NREL’s EVI-X model. A summary of these platforms and how NGP has worked to improve their effectiveness has also been added to this years’ report.

Thurston, Sam↗

Automation of Vulnerability and Patch Management: Information Extraction, Association, and Optimization

Vulnerability and patch management is an integral part of a robust cybersecurity program, yet it grows increasingly complex due to the sheer amount of data that must be analyzed. Particularly in Operational Technology (OT) environments, analysis must be done manually because of the lack of automated solutions. Additionally, there are many steps in this process, from the initial discovery of the vulnerability to the implementation of its remediation, and each step in the process requires different data in order to be performed effectively. In this work, we provide approaches and strategies to assist operators in industrial or OT environments throughout the vulnerability management cycle. Security advisories provide key information about mitigation strategies, or actions that can be taken when a patch is unavailable or cannot be installed. Details of these strategies are not shared in public vulnerability databases and must be found manually. We approach this problem by designing a solution to automatically identify that information within vendor security advisories and retrieve it for operator use. We start with an approach that requires domain-specific knowledge of certain frequently-seen reference websites. Next, an approach that can work on an arbitrary website but relies on certain keywords. Finally, an approach that uses Natural Language Processing (NLP) methods and does not require specific knowledge or keywords. Each of these approaches is more general than its predecessor; we demonstrate high accuracy for all approaches Advisories also often contain details of affected products in non-standard or natural language formats. While this information can be easily understood when read by an operator, the non-standard format acts as a barrier to effective automation. We provide an approach for the first step in this process: identifying vendors in security advisories and mapping them to a standard framework for representing digital assets and software products. We evaluate five established string similarity algorithms, plus one of our own design that combines string similarity and information theory, on the task of mapping vendors to their corresponding entries in the Common Platform Enumeration (CPE) repository. Our results show that our proposed metric outperforms all others. Due to the constraints on time, finances, and personnel for organizations, Large Language Models (LLMs) may seem like attractive opportunities for security operators to speed up information gathering; however, it is still not clear whether LLMs can handle vulnerability management tasks well. To answer this question, we perform an empirical study of LLMs’ ability to provide consistent, accurate information about vulnerabilities in order to guide organizations in their adoption of LLMs. We observe poor performance for all models tested, suggesting that these models are not well-suited to the consistent retrieval of accurate vulnerability information. Finally, once vulnerabilities have been identified and any additional information has been obtained, operators must decide which remediation actions to implement based on their available resources. This already-complex problem becomes even more so when we consider that a vulnerability may have multiple avenues for remediation. We formulate this scenario as two knapsack problems and provide solutions, which we then compare against several existing strategies for vulnerability prioritization seen in real operational environments.

McClanahan, Kylie↗

EV Profile Capture

NextGen Profiles' EV profile capture efforts aimed to explore the variance in performance and evaluate how different operational conditions influence production EV charging behavior. Data were collected at a frequency of 10 Hz from both the EV and EVSE during each charge session. These charge session parameters were then entered into a time-series database for further analysis. The data were gathered under different operational conditions to examine the effects of various factors such as battery state of charge, battery temperature, vehicle condition, smart charge management, and EVSE limitations. The EV profile capture dataset includes extensive high-power charging data from 16 different EVs—comprising light-, medium-, and heavy-duty vehicles—along with EVSE from various suppliers. To protect confidentiality, the EV and EVSE metadata are anonymized, and the publicly released datasets are aggregated to 0.1-Hz frequency.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

EV Profile Capture 2025: Next-Gen Profiles Project Report

As part of the Next-Gen Profiles (NGP) project, the profile capture and analysis of production electric vehicles undergoing high-power charging (HPC) is conducted over a wide range of conditions to explore variance and performance. Charge session parameters are collected from both the electric vehicle (EV) and electric vehicle supply equipment (EVSE) at a rate of 10Hz and entered into a time-series database for analysis. These charge profiles are captured under nominal and off-nominal conditions, exploring the impact of starting battery state of charge (SOC), battery temperature, vehicle condition, smart charge management (SCM), EVSE limitations and charging adapter usage. Nominal conditions are defined as ideal conditions that should transfer the maximum allowable energy in the minimum possible amount of time. Nominal condition profiles are compared across EVs to characterize state-of-the-art EV charging performance against one another. Off-nominal condition profiles are compared against their nominal condition profile counterparts to highlight the variance across less desirable starting conditions within a single EV.

33 ADVANCED PROPULSION SYSTEMS↗

Vedizar Fingerprinter

SAND2025-03289O Vedizar Fingerprinter simplifies the process of identifying devices on a network by analyzing traffic data. It uses a unique library to recognize different devices, making it easier for users to understand what is happening on their networks. This software is ideal for IT and operational technology environments, helping organizations monitor their networks effectively. By saving results in a database, it allows for easy access and review of device information. Users can enhance their network security and optimize performance without needing specialized hardware or technical expertise. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Jacobellis, John [Sandia National Lab. (SNL-CA), L↗

1994 SEMCOG Household-Based Person Trip Survey

The Southeast Michigan Council of Governments contracted with the Applied Management & Planning Group to conduct a travel behavior survey of 7,361 households from March 28, 1994, to June 10, 1994. The primary purpose of this study was to provide the council with a new database of travel behavior to assist in updating the region's transportation models. This data collection includes demographic, socioeconomic, travel, and mobile source emissions models for projecting future patterns of development, travel, congestion, and air pollution. Information about household characteristics and travel was collected using a one-day, activity-focused diary and a separate household survey. In an activity diary, respondents recorded 65,535 trips during the assigned the day.

1Hz data↗

Report priority gaps in high temperature thermodynamic data (Interim Progress Report)

This interim progress report (Level 4 Milestone Number M4SF-26LL010203023) summarizes research conducted at Lawrence Livermore National Laboratory (LLNL) within the Argillite Host Rock Properties & Processes SF-26LL01020302. Our focus is to assess gaps in data availability and understanding for radionuclide thermodynamics within the context of a “hot repository” concept and expand SUPCRT-NE database development to address higher temperatures needed for a DPC DGR disposal concept. The database is intended to inform the argillite GDSA baseline model. The leading European thermochemical database (Thermochimie) is only applicable to temperatures below 80°C. Thus, a US effort to integrate and expand upon other international thermodynamics database efforts is needed, particular if a “hot repository” concept moves forward.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

Datum: A Scientific Metadata Catalog

The data catalog market is currently flooded with a myriad of different products, but none serve the scientific community well. There are cloud-native tools like Databricks, Snowflake,to on-premise solutions like Collibra and Datahub. The common failing of all these tools however, is their inability to serve the scientific data community directly. Most catalogs are targeted towards financial, health, or user data - not sensor or scientific domain data. They also prioritize integrations that often don’t exist or are just starting to be used in the scientific realm - all while ignoring common scientific tools and file types. Datum is a catalog which targets the scientific data directly, including the tools and networks in which those tools are used. We work with the producers and consumers of the data where they are, targeting cloud and on-premise with a focus on classified networks. Datum is an Erlang/Elixir application. Technical Features Note: The features listed below are still under development and may change, slightly, upon final delivery of the product. File Formats - Datum has the ability to read additional metadata and provides processing pipelines for the following file formats: Plain Text, PDF, LaTeX, HTML, Open Document Format (.odt), XML, CSV/TSV (and other standard delimiters), OpenDocument Database and Spreadsheets, Geo-Referenced TIFF, Common Data Format, HDF/HDF5, LabView TDMS, Excel, DeltaTables, Parquet, Apache Iceberg, Apache Hudi and many others. Metadata Collection - Scanners for the local and networked file systems and cloud storage providers. Network integration with common databases such as MSSQL and MySQL. User Plugin System - Users are able to provide either file processing, metadata extraction, or sampling plugins in the programming language of their choice. Authentication/Authorization -: OIDC integration, SCIM provisioning and EntraID integration out of the box. Full user and group management system with a “least privilege” operating mode. Governance - Customizable data governance platform; dictate and enforce required metadata, enforce data embargos, and enforce user agreements and NDAs before data access. Ability to create health checks on data, rejecting abandoned or poorly curated data and automatically removing it from the search index. Ability for users to submit corrections. Search - Semantic search is a first class citizen. No licenses to expensive, external software required. Integrated use of vectors and vector-based search allows for AI agent integration at all levels of operation. Metadata Model - Display and control data’s lineage and connections to other data and data directories. Data is modeled after a filesystem - an organization instantly recognizable and navigable by most any user. CLI and SDK - Ships with a Command Line Interface (CLI) tool and with a fully-featured Python SDK. This allows for rapid and programmatic use of Datum by every level of user. Minimal Infrastructure - Datum ships as a single executable file and can be run on any operating system and most CPU architectures. Datum has no reliance on external databases, search indexing tools, or other outside services - and it runs equally well on edge computing devices, cloud services, or in a clustered HPC environment.

darrington, john↗

Environmental Compliance through Inventory Management

This poster serves as an overview of the summer 2024 Environmental Program Department Internship projects. These include updating the air emissions inventory to appropriately calculate the previous year's emissions for all relevant sources at the Batavia site, as well as beginning a site-wide chemical inventory to eventually establish a database to track all non-household chemicals at the lab.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Do we have globally representative data to understand soil processes?

Understanding and modeling soils and soil organic matter (SOM) are central to a variety of human needs, from food production to ecosystem management. Soil data have been collected for over a century, but the global spatial and process representativeness of soil data remains unclear. We assessed the representativeness of currently available soil data that could be used to understand a variety of SOM processes. We used 16 open-source soil databases and data from over 281,000 unique locations globally, categorizing the databases into three main data types necessary to understand SOM processes: soil carbon stocks and fluxes, mechanistic drivers of these stocks and fluxes, and soil carbon gain or loss potential. We found that stock and driver data have extensive global coverage. However, data on soil carbon gain or loss potential, particularly data describing change in soils over time such as time series data, are severely limited in their global coverage. We conclude that while significant strides have been made in measuring soil carbon stocks and fluxes, and their drivers, we are limited in global data related to changes in soils over time. Our recommendations for soil data generators are to ensure precise metadata reporting and prioritizing sampling in underrepresented areas like tropical, arctic, mountainous, wetland and arid regions. We also encourage designing revisit schemes that explicitly support change detection and reporting multi-modal datasets that can aid in model development. Targeted measurement of low coverage soil data types and regions is necessary for a range of applications including current and future biogeochemical predictions, and their management and policy implications.

carbon fluxes↗

Optimizing Management of Persistent Data Structures in High-Performance Analytics

Large-scale data analytics workflows ingest massive input data into various data structures, including graphs and key-value datastores. These data structures undergo multiple transformations and computations and are typically reused in incremental and iterative analytics workflows. Persisting in-memory views of these data structures enables reusing them beyond the scope of a single program run while avoiding repetitive raw data ingestion overheads. Memory-mapped I/O enables persisting in-memory data structures without data serialization and deserialization overheads. However, memory-mapped I/O lacks the key feature of persisting consistent snapshots of these data structures for incremental ingestion and processing. The obstacles to efficient virtual memory snapshots using memory-mapped I/O include background writebacks outside the application’s control, and the significantly high storage footprint of such snapshots. To address these limitations, we present Privateer, a memory and storage management tool that enables storage-efficient virtual memory snapshotting while also optimizing snapshot I/O performance. Here, we integrated Privateer into Metall, a state-of-the-art persistent memory allocator for C++, and the Lightning Memory-Mapped Database (LMDB), a widely-used key-value datastore in data analytics and machine learning. Privateer optimized application performance by 1.22× when storing data structure snapshots to node-local storage, and up to 16.7× when storing snapshots to a parallel file system. Privateer also optimizes storage efficiency of incremental data structure snapshots by up to 11× using data deduplication and compression.

Computer science↗