Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “database management”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

Data Qualification Report: SRNL Glass Composition-Properties (ComPro) Database

The Savannah River National Laboratory Glass Composition-Properties (ComPro) database is an extensive database containing pertinent composition and durability data to support the accelerated clean-up mission at the Defense Waste Processing Facility. The activities described in this data qualification report were performed to support the information contained in the database. There were two objectives of the original data qualification process. The first objective was to review supporting documentation to determine if DOE/RW-0333P Quality Assurance Requirements and Description had been implemented during the original work. If the DOE/RW-0333P Quality Assurance Requirements and Description had not been directly implemented during the original work, the second objective was to determine if the controls that were used were adequate to meet the intent of the DOE/RW-0333P Quality Assurance Requirements and Description. The results of these two objectives and the activities performed to support these decisions are described in this document. An assessment of each dataset was made to determine if the data were RW-0333P Compliant, RW-0333P Equivalent or Non-RW-0333P Compliant. The original data qualification was performed in accordance with E7, Conduct of Engineering Manual, Procedure 3.70, Revision 4, Qualification of Data. The specific method that was used was Equivalent Controls as described in E7, 3.70. Revision 2 of this document adds supporting information for the RW-0333P Compliant datasets added to Revision 3 of the database.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

U.S. Pacific Coast Workshop Report on Preconstruction Research Recommendations (U.S. Offshore Wind Synthesis of Environmental Effects Research (SEER) Project)

In May 2022, the U.S. Offshore Wind Synthesis of Environmental Effects Research (SEER) project team hosted a stakeholder workshop focused on preconstruction (baseline) research needs for potential floating offshore wind (OSW) energy development on the U.S. Pacific Coast, including California, Oregon, and Washington. Prior to the workshop, the SEER team developed a set of initial synthesized research recommendations that were identified based on a review of relevant, publicly available resources and with advisory group input. The workshop covered three marine life breakout groups on subsequent days to discuss research recommendations related to 1) marine mammals and sea turtles, 2) fish and invertebrates, and 3) birds and bats. As part of the workshop, over a hundred participants from the public and private sectors provided feedback on various aspects of the initial research recommendations, including associated data and knowledge gaps, benefits/limitations of available methods and technologies, and technological advancements or infrastructure needed to address the recommendation. Approximately 1,000 total comments were received on the workshop MURAL boards and were synthesized in this report. Based on workshop feedback, SEER developed a final database of over 500 specific research recommendations based on more than 40 resources. In Fall 2022, the full database and a tool with updated synthesized research recommendations were disseminated on Tethys (https://tethys.pnnl.gov) to assist with informing future funding opportunities and research programming. There is a continued need to improve awareness of the potential environmental effects, monitoring technologies, and management strategies for floating OSW energy development on the U.S. Pacific Coast. Coordination of these activities will require the sustained involvement of multiple stakeholders from across sectors. Beyond the baseline considerations discussed in this workshop, future state-of-the-science activities should be planned to consider research needs across wind energy life cycle phases for all relevant wildlife taxa and associated habitat and ecosystem processes.

17 WIND ENERGY↗

N 2 Onet: a global collaborative network facilitating advances in measurement, modeling, and mitigation of agricultural soil nitrous oxide emissions

Nitrogen (N) fertilizer supports global food production, but its use and overuse drive emissions of nitrous oxide (N 2 O), a potent and long-lived greenhouse gas. Understanding the drivers of N 2 O fluxes remains elusive, making it difficult to predict emissions in time and space and to develop and evaluate ways to lower emissions through management. Major scientific uncertainties underlying the understanding of the drivers of N 2 O fluxes identified in a workshop of N 2 O emissions experts include poor process-based understanding of controls on soil N 2 O emissions in the field; insufficient data to reduce uncertainty in N 2 O budgets from the field to regional scales, including N 2 O emission measurements and importantly, field-scale N balances; and high uncertainty in model predictions of soil N 2 O emissions across environmental and management conditions. To reduce these uncertainties, we present the concept of N 2 Onet, a global collaborative initiative to accelerate advances in N 2 O measurement, analyses, and mitigation. N 2 Onet will serve as an observational network of supersites with multi-scale measurements; a database hub for N 2 O flux and ancillary data; and a catalyst for community building, information sharing, and training. By coalescing and coordinating the global community of researchers, N 2 Onet will provide a roadmap for reducing N 2 O emissions from agriculture worldwide.

54 ENVIRONMENTAL SCIENCES↗

Myna

The additive manufacturing (AM) community has been developing digital factory tools over the past decade to better leverage the multi-modal process data coming out of the advanced manufacturing process. As a result, numerous databases of additive manufacturing process data exist in the literature and in the archival storage of disparate research groups. While some efforts have been made to create a standard ontology for storing and sharing AM data, in practice a variety of data structures are used to store AM build data, even within a single institution. This causes many problems for maintainability and extensibility when attempting to integrate computational modeling tools with experimental data to either validate models or to provide further insight into results and trends. Myna is a Python-based framework that aims to decrease the effort needed to connect individual computational models to the variety of AM process data that exist in different research groups and institutions. This type of software is sometimes referred to as "middleware" or “glueware,” in that it connects disparate databases and applications into a single computational ecosystem. Instead of maintaining unique interfaces between each application and each database, developers can create a single interface from each application to Myna and thereby gain access to the implemented database connections. Similarly, developing a database connection in Myna provides access to the developed simulation applications. This framework greatly simplifies the maintainability of model applications that rely on experimental data. Using external simulation tools, users will also be able to run pre-configured workflows using the built-in workflow manager. Several examples of input files are provided with Myna for different workflows, including melt pool geometry predictions and detailed melt pool and solidification microstructure predictions.

Knapp, GerryL. [Oak Ridge National Laboratory (ORN↗

DiMER

SAND2025-04145O DiMER is a Python based tool that helps researchers understand the functions of genes by searching through multiple biological databases. It takes user-provided data and scans various databases to find the best matches for gene functions, generating a clear summary of results. DiMER identifies the most relevant functional annotations and improves upon previous annotations by replacing instances of "unknown protein function" with more accurate descriptions. DiMER requires minimal setup. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Mageeney, Catherine [Sandia National Lab. (SNL-CA)↗

Digital Safety Analysis for Small Modular Nuclear Reactors (SMRs)

A Documented Safety Analysis (DSA) is a Department of Energy (DOE) construct that defines the extent to which a nuclear facility can be operated safely. It includes a description of hazards, safe boundaries, and hazard controls. The authors assert that a Digital Safety Analysis (DgSA) is far superior to a legacy DSA for several reasons: • The underling database is structured such that it is possible to perform a comprehensive design review and safety analysis by iterating systematically across a hierarchy of linked objects versus a redundant and spotty review by entities of various abilities under unknown resource and schedule constraints. • The analysis of a new design can discover elements that are similar to elements in previous designs. The discovery of similarities is made possible by using the same structure for the underlying database for each new DgSA. The “prior learning” from previous designs is then applied automatically to new designs. • Outputs from the DgSA are from a single source to ensure consistency among various views of the same information. After the DgSA is released, the continued use of a single source implements a configuration management program to ensure consistency between the design basis, the design, the built system, and system procedures. • The development of the DgSA is agile in that any change in a linked object triggers an analysis of impacts on other linked objects and updates of linked objects are made accordingly. After the DgSA is released, the continued maintenance of these links and objects automates the “unreviewed safety question” process.

22 - GENERAL STUDIES OF NUCLEAR REACTORS↗

Evaluation of New Additions to OLI Software in Predicting Mercuric and Mercurous Species in Liquid Waste Operations

Speciation of mercury during the pretreatment steps of tank waste processing is critical to successful mercury removal prior to vitrification during Liquid Waste Operations (LWO) at SRS. OLI software has been used to predict mercury speciation and activity throughout LWO. The OLI software operates based on a thermodynamic framework called the Mixed Solvent Electrolyte (MSE) framework. The MSE framework allows prediction in theoretically infinitely dilute to concentrated mixtures (e.g., purely solute solutions). Before modification to the MSE framework databanks, certain critical mercury species were missing in the MSE databank, and some thermodynamic data needed to be updated for the OLI software to accurately predict mercury chemical species in SRS waste tanks. To better reflect streams across LWO, new mercury species were integrated into the MSE database. To evaluate the changes to the OLI MSE framework per the Technical Task Request (TTR) and the Task Technical and Quality Assurance Plan (TTQAP), waste stream compositions from Tanks 38, 43, and Tank 50 decontaminated salt solution (DSS) were used as model inputs. Models were developed and executed using both the old and new databases. Compositional analyses from caustic Tank 50 DSS and caustic Tanks 38 and 43 were used as the input streams. These streams represent the most comprehensive chemical data sets where both mercury and tank constituents were measured together. Results for Tank 50 DSS predict HgO as the predominant species in both databases. Both methyl and dimethyl Hg species are present when the new database is ‘on’ and are not predicted with the new database turned ‘off’. The new database predicts a greater amount of HgO and a greater fraction of it in the solid phase. Pourbaix diagrams (potential vs. pH) generated for each Tank 50 DSS were identical regardless of which database was used. Elemental Hg and HgO were predicted in the water stable region under basic conditions. Tanks 38 and 43 follow similar trends as the Tank 50 DSS models. Unlike Tanks 38 and 50 DSS, the Tank 43 Pourbaix plot shows a region of stability for an aqueous HgOHCO3 - species between approximately pH 7-11. In all streams, when MeHg+ is included in the inputs, the new database predicts aqueous MeHgOH as the dominant species. If elemental or dimethyl mercury is in the waste stream, the new database model predicts they are unchanged and remain in those states and quantities. Additionally, the total mercury values are reported for both the measured input data and the OLI output data for all considered tanks. The summary indicates that the percentage error between the measured and calculated values is less than 1% in all cases The reconciliations and generation of the Pourbaix diagrams for Tank 50 DSS took approximately ten times longer with the new database ‘on’. In addition, over the course of that time, models with the new database ‘on’ were more likely to crash or display an error. Some modest performance improvements were noted when modeling with an i7 processor versus an i5. An example error is found in Appendix A. Furthermore, Appendix B provides V&V for two chemical systems analyzed with the OLI software, results were satisfactory. It is recommended to utilize the new databases (i.e., HCO.ddb and SR-Hg.ddb) in future Savannah River Mission Completion applications of OLI to represent pseudo steady-state. Furthermore, the integration and utilization of the new databases (i.e., HCO.ddb and SR-Hg.ddb) in modeling applications (e.g., Aspen) is also recommended.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

Locating Undocumented Wells Using Historical Oil and Gas Exploration Maps: A Case Study in Osage County, Oklahoma

Undocumented oil and gas wells lack reliable information about their locations and characteristics, making them difficult to identify. These wells can result in unanticipated delays and costs in the development of nearby surface and subsurface resources, and, if improperly plugged, can cause contamination. This study leverages historical petroleum exploration maps to locate such wells, focusing on Osage County, Oklahoma. Two sets of early 20th century oil and gas exploration maps by the United States Geological Survey were georeferenced and analyzed using a computer vision model to detect well symbols. The locations of detected wells were compared to the location of known wells in the database from the Bureau of Indian Affairs Osage Agency to identify potential undocumented wells. The analysis yielded over 500 potential undocumented wells, with dry holes constituting the largest fraction. Field verification confirmed the presence of some undocumented wells. Comparison with prior work revealed limited overlap, underscoring the complementary value of historical oil and gas maps for locating undocumented wells. This approach demonstrates the utility of integrating historical cartographic resources with modern geospatial and machine learning techniques to improve the identification and management of undocumented wells.

Energy - Petroleum↗

An Overview of the Molten Salt Thermal Properties Database--Thermophysical, Version 3.1 (MSTDB-TP v.3.1)

This report presents the current status of the Molten Salt Thermal Properties Database–Thermophysical (MSTDB-TP). Information regarding version 3.1 is provided herein, which contains 820 individual salt entries (data from 180+ independent studies); the thermophysical properties contained in the database include density, viscosity, thermal conductivity, and heat capacity. The major updates to the database include a significant expansion of pseudobinary and higher-order chloride salt mixtures, many of which bearing actinides, and an incorporation of more recent literature data (i.e., that within the past 5 years). Also, modifications have been made to the pure compound data in the database as a consequence of an external quality assessment of duplicate datasets. The user-facing API for the MSTDB-TP, Saline, has been updated to include viscosity estimation capabilities based on the Redlich–Kister formalism; this is an advancement with respect to the existing density estimation capabilities. The graphical user interface was also updated to include a density estimation capability, backed by Saline. Finally, additional preliminary efforts to include surface tension into the database, as well as an investigation on formalisms that would be appropriate for thermal conductivity estimation, are reported herein.

36 MATERIALS SCIENCE↗

Public water supply infrastructure extensification and diversification in surface waters is insufficient to meet future demands in Texas

The data were developed to evaluate the capacity of existing and potential new surface water supply infrastructure to meet projected public water demands across districts in Texas under multiple future socioeconomic and climate scenarios. The database integrates hydrologic, water quality, infrastructure, energy, cost, demographic, and demand-projection information for candidate surface water supply locations. Candidate sites include stream reaches, waterbodies, reservoir surplus locations, and potential new reservoir sites. Water availability is characterized using historical and projected flow conditions, while site suitability is evaluated using five indicators: Water Availability Index (WAI), Water Quality Index (WQI), Energy Requirement Index (ERI), Water Treatment Cost (WTC), and Water Infrastructure Cost (WIC). The datasets include statewide candidate-site information, district-level demand projections under Shared Socioeconomic Pathways (SSPs), runoff-based allocation constraints, climate-stress metrics, and optimization outputs evaluating alternative infrastructure planning strategies. Optimization results compare Business-as-Usual (BAU) and All Surface Water (AllSW) demand-management approaches under both scaled and fixed cost-cap strategies. Associated validation datasets provide district-level feasibility assessments, infrastructure selection outcomes, cost-cap utilization, demand satisfaction metrics, and constraint diagnostics. Additional datasets quantify projected changes in storage and flow conditions as well as water availability stress for both existing and newly selected intake locations under the SSP5 scenario for mid-century and late-century climate conditions. Together, these datasets support assessment of the extent to which surface-water infrastructure expansion and diversification strategies can satisfy future public water demands while accounting for hydrologic, economic, and planning constraints across Texas. Dataset(s) Description Dataset_preoptimization.xlsx Comprehensive pre-optimization dataset containing candidate water-supply sites and associated hydrologic, water-quality, infrastructure, climate, demographic, runoff, and demand-projection variables used as inputs to the optimization analyses. Includes variable descriptions and the full statewide candidate-site database. District_level_site_selection.zip - Compressed archive containing all SSP-specific district-level optimization result files MESIO_ssp1_results.xlsx District-level site selection results for SSP1 (MESIO). Includes variable descriptions, BAU and AllSW site-selection results under scaled and fixed cost strategies, and district-level validation diagnostics. MESID_ssp2_results.xlsx District-level site selection results for SSP2 (MESID). Includes variable descriptions, BAU and AllSW site-selection results under scaled and fixed cost strategies, and district-level validation diagnostics. LCMRD_ssp3_results.xlsx District-level site selection results for SSP3 (LCMRD). Includes variable descriptions, BAU and AllSW site-selection results under scaled and fixed cost strategies, and district-level validation diagnostics. IRDev-Low_ssp4l_results.xlsx District-level site selection results for SSP4-Low (IRDev-Low). Includes variable descriptions, BAU and AllSW site-selection results under scaled and fixed cost strategies, and district-level validation diagnostics. IRDev-High_ssp4h_results.xlsx District-level site selection results for SSP4-High (IRDev-High). Includes variable descriptions, BAU and AllSW site-selection results under scaled and fixed cost strategies, and district-level validation diagnostics. RSIM_ssp5_results.xlsx District-level site selection results for SSP5 (RSIM). Includes variable descriptions, BAU and AllSW site-selection results under scaled and fixed cost strategies, and district-level validation diagnostics. tx_hydrological_stress.xlsx Hydrological stress dataset for existing and newly selected intake locations. Includes projected mid-century and late-century changes, gain/loss classifications, planning strategy information, and accompanying variable descriptions. Also includes water-stress metrics derived from historical and projected low-flow conditions.

Okoye, Perpetua I. (ORCID:0000000215545033)↗

DICOMs, Missiles, and Metadata: The U.S. Nuclear Weapons Program Leverages a Medical Standard

Digital Imagine and Communications in Medicine (DICOM) images, although most commonly used in medical settings, have been widely adopted by the United States Department of Energy (DOE) for capturing images of the internal components in a nuclear weapon. DICOMs, a lesser-known file format that combine nested metadata structures with complex image data-including multiple planes, frames, and high resolution-require the creation of access copies to support usability within the DOE. Used primarily for ensuring the safety, security, and reliability of the U.S. nuclear stockpile, the Los Alamos National Laboratory (LANL)’s DICOM images and corresponding image metadata must be accessible to scientists and researchers via our institutional centralized databases. This paper describes the author's creation of a Python script that converts DICOM images into accessible, archive-friendly TIFF files while preserving key image data and metadata.

96 KNOWLEDGE MANAGEMENT AND PRESERVATION↗

Mining Thermophile Photosynthesis Genes: A Synthetic Operon Expressing Chloroflexota Species Reaction Center Genes in Rhodobacter sphaeroides

Photosynthesis is the foundation of the vast majority of life systems, and is therefore the most important bioenergetic process on earth. The greatest diversity of photosynthetic systems is found in microorganisms. However, our understanding of the biophysical and biochemical processes that transduce light into chemical energy is derived from a relatively small subset of proteins from microbes that are amenable to cultivation, in contrast to the huge number of predicted proteins that catalyze the initial photochemical reactions deposited in databases, such as from metagenomics. We describe the use of a Rhodobacter sphaeroides laboratory strain for the expression of heterologous photosynthesis genes to demonstrate the feasibility of mining this resource, focusing on hot spring Chloroflexota gene sequences. Using a synthetic operon of genes, we produced a photochemically active complex of reaction center proteins in our biological system. We also present bioinformatic analyses of anoxygenic type II reaction center sequences from metagenomic samples collected from hot (42–90 °C) springs available through the JGI IMG database, to generate a resource of diverse sequences that are potentially adapted to photosynthesis at such temperatures. These data provide a view into the natural diversity of anoxygenic photosynthesis, through a lens focused on high-temperature environments. The approach we took to express such genes can be applied for potential biotechnology purposes as well as for studies of fundamental catalytic properties of these heretofore inaccessible protein complexes.

Chloroflexota↗

An open-access simulated earthquake ground-motion database for an M7 Hayward Fault earthquake in the San Francisco Bay Region

Comprehensive understanding of earthquake ground motions, particularly in the near-fault region of large-magnitude events, is limited by gaps in strong-motion data. This challenge is prominent in areas with high seismic hazard but infrequent large earthquakes where data is sparse and difficult to interpret. These data limitations lead to uncertainties in the development of site-specific ground motions, which are crucial for engineering risk assessments. To address these challenges, physics-based regional-scale ground-motion simulations have been developed. With the emergence of exaflop-scale computing ecosystems, it is now possible to simulate regional earthquake processes at unprecedented fidelity and generate the large number of fault rupture realizations necessary to characterize both intra- and inter-event ground-motion variability. This article introduces a new database of simulated earthquake ground motions, created for applications in earthquake engineering, earthquake planning, and emergency response. The inaugural version of the database features simulated ground motions for a magnitude 7 Hayward Fault earthquake in the San Francisco Bay Region (SFBR), using the EarthQuake SIMulation (EQSIM) simulation framework and the Graves–Pitarka kinematic rupture model. The aim is to provide high-fidelity, spatially dense, three-component motions generated on the Department of Energy’s (DOE) newest generation of graphics processing unit (GPU)-accelerated supercomputers. These motions are being made openly available to the engineering, scientific, and disaster planning communities. In addition, this work develops protocols for the efficient dissemination of these large data sets and emphasizes community engagement to build confidence in their application. This article discusses the methodology behind the data, underlying software verification and validation, scalable data management, and a user interface for data access. The goal is to facilitate widespread use and elicit expert feedback to maximize the utility and exploitation of simulated motions. While the initial focus is on the San Francisco Region, simulations for additional regions will be added as the DOE program progresses.

Simulated ground-motion database↗

DUNE Rucio Server Scalabiilty Studies

The DUNE collaboration has an ongoing production effort to simulate the full detectors and to analyze the various prototypes that are currently running. Rucio is used to manage the 40PB of files made to date. When 500 or more jobs were sending output to Rucio simultaneously via Rucio upload, we observed timeouts, unhandled exceptions, and Rucio server restarts due to slow performance. In collaboration with the core Rucio team we did a full review of the Rucio upload code and identified several optimizations that can be made. We also have deployed the Ingress load balancer in front of our Rucio servers and added a database connection pooling utility. These changes led to significant improvement both in reliability and scalability, yet we anticipate even better performance will eventually be required. We describe in this paper the initial state of the system, the various debugging processes that were used, and our plans to further improve scalability.

Calcutt, J. [Brookhaven Natl. Lab.]↗

American-Made Solar Prize: Edgeli Enables DER Integration (CRADA 615) (Final Report)

The purpose of this project was to demonstrate how granular time series data and automated data transformation, and impact assessment tools could speed interconnection approvals for distributed energy resource projects of various types and sizes. Types included community solar, rooftop solar, and EV charging projects. Using software routines to automate the transformation of data (e.g. GIS) to a network database and power flow model then applying scenarios to create hourly (8760) hosting capacity values and voltage and thermal impacts for specific projects, we were able to demonstrate the feasibility of quickly assembling and analyzing key utility data sets for interconnection purposes. The outcomes of this effort will become the foundation for future work that will enhance and encapsulate the software components developed as part of this project, into web services (e.g. APIs) that can be integrated into queue management systems and automate interconnection screening processes.

14 SOLAR ENERGY↗

Integrase-on-Demand

SAND2025-07449O Integrase-on-Demand is a software tool that allows users to identify regions in genomic sequences where genetic material can be integrated with high probability. It uses a database of integrases and their DNA attachment sites to search against any genomic sequence, producing a list of open sites, the integrase sequence, and the source of the genomic island. The program requires MASH software to be available on the system. It consists of a main script and a precomputed input file, with a taxonomy mode that searches closely related genomes and a search mode that looks for identical attachment site matches in the integrase/attachment input file. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Williams, Kelly [Sandia National Lab. (SNL-CA), Li↗

Site Characterization of the Highest-Priority Geologic Formations for CO2 Storage in Wyoming

The project Site Characterization of the Highest-Priority Geologic Formations for CO2 Storage in Wyoming is one of 9 site characterization projects that were implemented as part of ARRA (American Recovery and Reinvestment Act). Data from this project was used to improve resolution of data in NATCARB in the area of study. Data related to this study has already been incorporated in NATCARB Atlas. The Wyoming Carbon Underground Storage Project (WY-CUSP) consisted of CO2 storage site characterization and evaluation, focusing on Wyoming’s most promising CO2 storage reservoirs (the Pennsylvanian Weber/Tensleep Sandstone and Mississippian Madison Limestone) and premier CO2 storage site (Rock Springs Uplift). Results from the WY-CUSP project suggest the two reservoirs could store up to 17,000 million tons of CO2. The WY-CUSP team drilled a stratigraphic test well and acquired a 3-D seismic survey covering 25 square miles of the Rock Springs Uplift site. The team retrieved 916 feet of core from the 12,810-foot-deep well, along with a complete log suite, borehole images, fluid samples, and other data. Project partners (1) provided continuous visual documentation of the core, including grain size, mineralogy, facies distribution, and porosity; (2) performed continuous permeability and velocity scans of selected reservoir intervals; and (3) chemically analyzed the fluid samples. WY-CUSP scientists integrated seismic attributes with observations from log suites, a VSP survey, core, fluid samples, and laboratory analyses, including continuous permeability scans. From these integrations, researchers constructed 3-D spatial distribution volumes of reservoir and seal properties that represent geological heterogeneity at the targeted CO2 storage site. The WY-CUSP team used this data to perform new CO2 plume migration simulations. Baker Hughes, Inc., completed a series of small-scale, in-situ water injectivity measurements. A database was formed when observations, analyses, and experiments from the stratigraphic test well were integrated. Correlation of these data allowed petrophysical parameters to be extrapolated from the test well out into the storage domain (5x5 mile 3-D seismic survey volume). This resulted in an improved, realistic understanding of performance assessments for potential CO2 storage scenarios. The WY-CUSP team worked on (1) improving CO2 storage resource estimates, (2) establishing long-term integrity and permanence of confining layers, (3) designing a profitable strategy for pressure management, and (4) evaluating the utilization of stored CO2 at the Rock Spring Uplift. Finally, Baker Hughes developed a microseismic baseline for the test site using in-bore geophones to complete field operations.

3-D seismic↗

Metagenome-assembled genomes from Wind River Basin floodplain sediments Riverton, Wyoming site (May to September 2017)

Microorganisms play a key role in cycling nutrients and contaminants in the terrestrial environment depending on their genetic potential. Here we present metagenome-assembled genomes (MAGs) for the bacterial and archaeal community in floodplain sediment samples taken roughly every month in the period May 18 to September 13 in 2017 at a location (Pit2) close to DOE Legacy Management well 855 at the Riverton, Wyoming floodplain site in the Wind River Basin (WRB). The groundwater at this site exhibits persistent U, Mo, and sulfate plumes and is one of the field sites in focus for the SLAC Groundwater Quality SFA program. Cores were taken with a hand-auger and separated into 5-20 cm segments based on soil horizonation down to 150 cm depth below surface. Each segment was subsampled for microbial analyses. Corresponding 16S rRNA gene amplicon data is available at the NCBI Single Read Archive (SRA) Database BioProject ID PRJNA626616, and soil geochemistry data at doi:10.15485/1631972. 40 metagenomes were sequenced through JGI and can be found under Gold sequencing project: Gs0142591. Metagenomes were assembled, binned, and refined using metawrap to generate MAGs (>50% complete and < 10% contamination based on checkM scores). This dataset includes a zip file of 6993 MAG fasta files and a csv file with quality, taxonomic classification (GTDB RS220), and metagenome accessions for MAGs generated from the Wind River Basin (WRB). This dataset also includes a file-level metadata (flmd.csv) file that lists each file contained in the dataset with associated metadata and a data dictionary (dd.csv) file that contains column/row headers used throughout the files along with a definition, units, and data type.

54 ENVIRONMENTAL SCIENCES↗