Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “data repository”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Data for publication: "A fresh take: Seasonal changes in terrestrial freshwater inputs impact salt marsh hydrology and vegetation dynamics"

This data repository contains data associated with the manuscript "A fresh take: Seasonal changes in terrestrial freshwater inputs impact salt marsh hydrology and vegetation dynamics". This study was conducted at the Elkhorn Slough National Estuarine Research Reserve in Watsonville, California from October 2019 - June 2022. We sought to understand the role of shallow freshwater inputs from adjacent uplands on salt marsh hydrologic behavior and vegetation productivity. This dataset contains CSV files of the following: daily salt marsh subsurface water level and pore water conductivity, monthly vegetation survey measurements, soil core data. Estuary surface water level, conductivity, and local precipitation were downloaded from the National Estuarine Research Reserve System (Centralized Data Management Office) at https://cdmo.baruch.sc.edu/.

54 ENVIRONMENTAL SCIENCES↗

Seismology in the Cloud: A New Streaming Workflow

Data-intensive research in seismology is experiencing a recent boom, driven in part by large volumes of available data and advances in the growing field of data science. However, there are significant barriers to processing large data volumes, such as long retrieval times from data repositories, complex data management, and limited computational resources. New tools and platforms have reduced the barriers to entry for scientific cluster computing, including the maturation of the commercial cloud as an accessible instrument for research. Here, we build a customized research cluster in the cloud to test a new workflow for large-scale seismic analysis, in which data are processed as a stream (retrieved on-the-fly and acted upon without storing), with data from the Incorporated Research Institutions for Seismology Data Management Center. We use this workflow to deploy a spectral peak detection algorithm over 5.6 TB of compressed continuous seismic data from 2074 stations of the USArray Transportable Array EarthScope network. Using a 50-node cluster in the cloud, we completed the noise survey in 80 hr, with an average data throughput of 1.7 GB per minute. By varying cluster sizes, we find the scaling of our analysis to be sublinear, due to a combination of algorithmic limitations and data center response times. The cloud-based streaming workflow represents an order-of-magnitude increase in acquisition and processing speed compared to a traditional download-store-process workflow, and offers the additional benefits of employing a flexible, accessible, and widely used computing architecture. It is limited, however, due to its reliance on Internet transfer speeds and data center service capacity, and may not work well for repeated analyses or those for which even higher data throughputs are needed. These research applications will require a new class of cloud-native approaches in which both data and analysis are in the cloud.

58 GEOSCIENCES↗

Integrated GW Farm ABM

This Data Repository includes data used for the integrated groundwater- farm ABM model, raw model output from scenario ensemble, and processed outputs that isolate the groundwater storage depletion outcomes for the 35,000 farm cells. Model Inputs: Farm ABM Inputs: This folder contains the input data used by the integrated groundwater - farm ABM modelling script (Python file) used for the high performance computing (HPC) experiments. The sub-folder "data inputs" contains all of the farm attribute data, while the three files in the folder have the hydrogeological data lookup table (NLDAS Cost Curve Attributes.csv), a lookup table (Theis well function table.csv) for the groundwater cost curve function, and the farm indexes and corresponding NLDAS ids for all of the cells run in this experiment (nldas farms subset final.csv). NLDAS Cost curve hydrogeological data: Hydrogeological data aggregated to 1/8 degree resolution and aligned with the NLDAS grid. Parameters include: water depth below ground surface [meters], subsurface porosity [unitless], aquifer depth from ground surface to aquifer bottom [meters], annual average recharge (USGS: mm, Doll: meters), and three different hydraulic conductivity (K) values (meters/day). The three K values represent the mean value from Gleeson et al. (2018), one standard deviation above the mean from Gleeson et al. (2018), and the de Graaf et al. 2020 modifications to certain lithologies. Additional information about these datasets and their processing are documented in the supplement to Yoon et al. 2025 (in review). Output: Raw outputs: This folder contains a .zip file that has model outputs for the entire scenario ensemble. There is one csv for each farm id, using the format "farm farmid cases.csv". The relationship between the farm id and NLDAS id is defined by the "nldas farms subset final.csv" located in the Farm ABM Inputs folder. Each csv has 625 rows, corresponding to 625 combinations of different scenario parameter values. Each row (scenario) represents the outcome of a 100 year simulation. Columns define scenario settings and summary statistics for each scenario. The first four columns define the scenario settings: "hydro ratio," "econ ratio," "K scenario," and "gamma scenario." The hydro and econ ratios are values passed to the modeling script that influence multipliers for other model parameters, as documented in the supplement to Yoon et al. 2025 (in review). The gamma multiplier is a coefficient multiplier applied to the baseline gamma values (values below 1 represent lower unobserved costs compared to baseline, values above 1 represent higher costs). The K scenario names represent K values of: "low": 0.5 m/d, "int 1": 2.5 m/d, "int 2": 10 m/d, "high": 50 m/d, and "gleeson": mean Gleeson K value. "Perc vol depleted" is the fraction of groundwater depleted at the end of the 100 simulation. Processed Output: Derived depletion outcomes from raw outputs: All of the individual csv files from the Raw outputs were aggregated into a single file that has the scenario settings and fraction depletion "Perc vol depleted" for every farm cell, for every scenario. The other two files define relationships between the farm id, NLDAS id, and local and major aquifer units, used for aquifer-level depletion analysis.

Agent based modeling↗

MSD CoP Webinar: "Advances in MSD-LIVE to Support the MSD Community of Practice"

Context: This webinar was hosted by the MultiSector Dynamics Community of Practice (MSD CoP; https://multisectordynamics.org). Advances in MSD-LIVE to Support the MSD Community of Practice Presenters: Casey Burleyson and Zoe Guillen (Pacific Northwest National Laboratory) Abstract: The MultiSector Dynamics Living, Intuitive, Value-adding, Environment (MSD-LIVE; msdlive.org) is a cloud-based data management system and advanced computing platform that enables MSD researchers to document and archive their data, run their models and analysis tools, and share their data, software, and workflows within the MSD Community of Practice. Recently, several high-profile datasets have attracted many new users to MSD-LIVE. This webinar has two goals: 1) To refamiliarize the MSD community and new users with the components of the platform (e.g., the data repository, model training notebooks, and data dashboards) and to highlight examples of how these components are advancing MSD science and 2) To demonstrate new features in v3 of the platform, released in late 2025. The main new feature in v3 is the ability to interactively explore data in MSD-LIVE without downloading it. MSD-LIVE users can now click a button in our data repository and launch a blank Jupyter notebook with access to the underlying data on AWS. Users can use the notebook to write analysis, visualization, or subsetting routines that process the data directly on the AWS cloud. We also added a GitHub integration feature that allows users to share analysis or visualization code they develop with the community of MSD-LIVE users. The webinar will wrap up with a look at what's coming next for MSD-LIVE in 2026. Moderator: Patrick M. Reed (MSD CoP Facilitation Team) This webinar was held on: May 12th, 2026 from 1-2 PM EST.

Open Science↗

Complete β-decay patterns of 142 Cs, 142 Ba, and 142 La determined using total absorption spectroscopy

Background: The β decays of fission products produced in nuclear fuel are important for nuclear energy applications and fundamental science of reactor antineutrinos. In particular, nuclear reactor safety is related to the decay modes of radioactive neutron-rich nuclei, primarily via the emission of γ rays, neutrons, and electrons. Additionally nuclear reactors are the most powerful man-made source of antineutrinos emitted during the β decay of fission products. These antineutrinos are used to inspect fundamental properties of leptons as well as informing reactor operation. However, the majority of data on complex decays of fission products collected in the evaluated nuclear data repositories like Evaluated Nuclear Structure Data File (ENSDF) and Evaluated Nuclear Data Files (ENDF) are based on low-efficiency and often incomplete measurements resulting in questionable reference reactor antineutrino flux predictions, see the analysis by [Nichols, J. Nucl. Sci. Technol. 52, 17 (2015)]. Various assessments like the one done under the auspices of the [Yoshida et al., Assessment of Fission Product Decay Data for Decay Heat Calculations: A report by the Working Party on International Evaluation Co-operation of the Nuclear Energy Agency Nuclear Science Committee (Nuclear Energy Agency, Organization for Economic Co-operation and Development, Paris, France, 2007), Vol. 25], as well as by [Sonzogni, Johnson, and McCutchan, Phys. Rev. C 91, 011301(R) (2015)] and [Dwyer and Langford, Phys. Rev. Lett. 114, 012502 (2015)], list the A = 142 isobars with high cumulative fission yield among the important nuclei where data for reactor decay heat and/or antineutrino production should be verified and/or improved. Purpose: Here, our goal is to improve the quality of β -decay measurements and evaluate the impact of modified decay schemes on reactor decay heat and antineutrino energy spectra, for fission products along the A = 142 isobaric chain. This work is an in depth follow-up on [Rasco et al., Phys. Rev. Lett. 117, 092501 (2016)]. which presented briefly the impact of the corrected decay scheme of 142 Cs . Here, we extend the data to full isobaric decay chain including the daughter nuclei, 142 Ba and 142 La, and present more details on the 142 Cs results. Method: The decays of neutron-rich isobars of mass A = 142 produced by means of proton-induced fission of 238 U were measured using the Modular Total Absorption Spectrometer (MTAS) array on-line at the mass separator and Tandem accelerator at Oak Ridge National Laboratory. Results: The β -decay schemes for 142 Cs and 142 La were modified with respect to the nuclear data repositories. A small β-delayed neutron branching ratio for 142 Cs emitter was remeasured as $0.10^{+5}_{–3}% %. Improved precision on the measured half-lives is reported. Small corrections to the low-energy decay of 142 Ba are made. The β-decay patterns for 142 La and 142 Cs are presented. The decay heat release and cross section for the detection of reactor antineutrinos are deduced and compared to earlier results. Conclusions: The β-feeding pattern for 142 Cs having decay energy value $Q_β$ of over 7 MeV was substantially modified with respect to the current ENSDF entry. Smaller changes were encountered for 142 La, but since this A = 142 isobar also has a large cumulative yield in fission, the changes influence both decay heat and the antineutrino spectra. The previously known β intensities for 142 Ba decay ($Q_β$ value of 2.2 MeV) were verified and slightly modified. Overall, increased decay heat values and lower flux of antineutrinos interacting with matter are presented.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

A cost and community perspective on the barriers to microbiome data reuse

Microbiome research is becoming a mature field with a wealth of data amassed from diverse ecosystems, yet the ability to fully leverage multi-omics data for reuse remains challenging. To provide a view into researchers’ behavior and attitudes towards data reuse, we surveyed over 700 microbiome researchers to evaluate data sharing and reuse challenges. We found that many researchers are impeded by difficulties with metadata records, challenges with processing and bioinformatics, and problems with data repository submissions. We also explored the cost constraints of data reuse at each step of the data reuse process to better understand “pain points” and to provide a more quantitative perspective from sixteen active researchers. The bioinformatics and data processing step was estimated to be the most time consuming, which aligns with some of the most frequently reported challenges from the community survey. From these two approaches, we present evidence-based recommendations for how to address data sharing and reuse challenges with concrete actions for future work.

59 BASIC BIOLOGICAL SCIENCES↗

Lost and Found: Rediscovering Microbiome-Associated Phenotypes that Reshape Agricultural Sustainability

Overview Code and data repository for NIL Manuscript. Documentation includes sequence processing examples and data analysis. Supplemental sequence processing and R statistical analysis for publication, which compares the microbiome of teosinte-B73 Near Isogenic Lines. Sample Data Amplicon sequence data for 16S rRNA genes, the fungal ITS2 region, and nitrogen-cycling functional genes are available through the NCBI Sequence Read Archive (SRA) under accession number PRJNA1042643(https://www.ncbi.nlm.nih.gov/bioproject/PRJNA1042643). Raw metabolomic data are available on Metabolomics Workbench, Project ID: PR002654. This study is available at the NIH Common Fund's National Metabolomics Data Repository (NMDR) website, the Metabolomics Workbench, https://www.metabolomicsworkbench.org where it has been assigned Study ID ST004211. The data can be accessed directly via its Project DOI: http://dx.doi.org/10.21228/M8KV8T.

Near Isogeneic Lines↗

Emerging materials intelligence ecosystems propelled by machine learning

We report that the age of cognitive computing and artificial intelligence (AI) is just dawning. Inspired by its successes and promises, several AI ecosystems are blossoming, many of them within the domain of materials science and engineering. These materials intelligence ecosystems are being shaped by several independent developments. Machine learning (ML) algorithms and extant materials data are utilized to create surrogate models of materials properties and performance predictions. Materials data repositories, which fuel such surrogate model development, are mushrooming. Automated data and knowledge capture from the literature (to populate data repositories) using natural language processing approaches is being explored. The design of materials that meet target property requirements and of synthesis steps to create target materials appear to be within reach, either by closed-loop active-learning strategies or by inverting the prediction pipeline using advanced generative algorithms. AI and ML concepts are also transforming the computational and physical laboratory infrastructural landscapes used to create materials data in the first place. Surrogate models that can outstrip physics-based simulations (on which they are trained) by several orders of magnitude in speed while preserving accuracy are being actively developed. Automation, autonomy and guided high-throughput techniques are imparting enormous efficiencies and eliminating redundancies in materials synthesis and characterization. The integration of the various parts of the burgeoning ML landscape may lead to materials-savvy digital assistants and to a human-machine partnership that could enable dramatic efficiencies, accelerated discoveries and increased productivity. Here, we review these emergent materials intelligence ecosystems and discuss the imminent challenges and opportunities. The materials research landscape is being transformed by the infusion of approaches based on machine learning. This Review discusses the emerging materials intelligence ecosystems and the potential of human-machine partnerships for fast and efficient virtual materials screening, development and discovery.

36 MATERIALS SCIENCE↗

ESS-DIVE Unoccupied Aerial Systems (UAS) Reporting Format v1

Here we present documentation of the ESS-DIVE reporting format for Unoccupied Aerial System (UAS) data and metadata. This reporting format provides guidance to data contributors on how to store data to maximize their discoverability, facilitate their efficient reuse, and add value to individual datasets. For data users, the reporting format will better allow data repositories to optimize data search and extraction, and more readily integrate similar data into harmonized synthesis products. The reporting format provides templates and guidance for the reporting of metadata for UAS experimental campaigns, individual flights, platform and sensor description. To improve data access and discoverability, the reporting format proposes a data description scheme of Levels based on the degree of processing, where Level 0 includes raw data, through to Level 3 being derived data end products. A range of examples of data types for each Level are given, with suggested file naming schemes. The reporting format presented here is intended to form a foundation for future development that will accommodate new UAS technologies and approaches to data access and use in the future. The reporting format documentation is maintained and updated on the ESS-DIVE Community Space GitHub at https://github.com/ess-dive-community/essdive-uas. This data package is the first published version of this reporting format, and comprises a zip file of the complete content of https://github.com/ess-dive-community/essdive-uas v1.0. The zip contains the reporting format description, instructions and variable definitions in GitHub markdown language (*.md) and metadata templates in csv format. The reporting format is designed to be compatible with other ESS-DIVE formats, and it is specifically recommended that this reporting format be used in conjunction with the File-level metadata (FLMD) and comma separated values (csv) reporting formats for submission to the ESS-DIVE repository.

54 ENVIRONMENTAL SCIENCES↗

Filling in Subsurface Storage Open Data Gaps - Updates to CCS Data Availability on EDX and EDX Spatial (FWP-1022465)

There is a need to preserve and efficiently access data resources to drive the next generation of research and development while ensuring compliance with DOE regulations. Over the last 10+ years, there has been ongoing efforts by the DOE Carbon Storage Program to ensure that there is effective data curation and preservation of DOE funded research leveraging the NETL-FECM data repository, the Energy Data eXchange (EDX). This talk presents updates about ongoing efforts to continue to support the mission of ensuring that carbon storage data is findable, accessible, interoperable, and reusable to the carbon storage stakeholder community through EDX and EDX Spatial. Presented at the NETL Carbon Management Review Meeting, Pittsburgh, 2024.

Morkner, Paige↗

ESS-DIVE Reporting Format for Dataset Package Metadata

ESS-DIVE’s (Environmental Systems Science Data Infrastructure for a Virtual Ecosystem) dataset metadata reporting format is intended to compile information about a dataset (e.g., title, description, funding sources) that can enable reuse of data submitted to the ESS-DIVE data repository. The files contained in this dataset include instructions (dataset_metadata_guide.md and README.md) that can be used to understand the types of metadata ESS-DIVE collects. The data dictionary (dd.csv) follows ESS-DIVE’s file-level metadata reporting format and includes brief descriptions about each element of the dataset metadata reporting format. This dataset also includes a terminology crosswalk (dataset_metadata_crosswalk.csv) that shows how ESS-DIVE’s metadata reporting format maps onto other existing metadata standards and reporting formats.Data contributors to ESS-DIVE can provide this metadata by manual entry using a web form or programmatically via ESS-DIVE’s API (Application Programming Interface). A metadata template (dataset_metadata_template.docx or dataset_metadata_template.pdf) can be used to collaboratively compile metadata before providing it to ESS-DIVE.Since being incorporated into ESS-DIVE’s data submission user interface, ESS-DIVE’s dataset metadata reporting format, has enabled features like automated metadata quality checks, and dissemination of ESS-DIVE datasets onto other data platforms including Google Dataset Search and DataCite.

54 ENVIRONMENTAL SCIENCES↗

Building partnerships for development of sustainable energy systems with atmospheric measurements

Atmospheric dynamics often play a critical role in the sustainability and reliability of diverse forms of energy production. This is especially true for the growing number of renewable energy deployments that harness aspects of the environment for power production. While the University of Memphis has a strong research background in energy systems, we have little experience working with the Earth and Environmental Systems Science Division (EESSD) and their associated User Facilities. Of particular interest to us is the Atmospheric Science Research and the Atmospheric Radiation Measurement (ARM) user facility to address surface-boundary layer interactions and physical phenomena. One of the major challenges for understanding and developing energy systems and management platforms is accurate modeling/forecasting of atmospheric conditions across disparate spatial and temporal scales. These conditions are often required to understand the lowest levels of the atmospheric boundary layer, but are also important to understand higher atmospheric conditions where aerosols affect cloud development. The objective of this work was to develop partnerships with national laboratories for collaboration on environmental science and its intersection with sustainable energy systems, as well as to leverage the ARM user facility data repositories to enhance our research capabilities in energy systems and their inter-dependence on environmental systems for future engagement with EESSD. Specifically, we accomplished these objectives by (1) developing collaborations with Oakridge National Laboratory ARM Data Science and Integration Group which resulted in student internships, (2) employed ARM data to develope modeling of the atmospheric boundary layer optical turbulence, and (3) optimally-sized large-scale renewable energy systems and their associated energy storage systems with ARM repository data.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Data and scripts associated with “Riverine dissolved organic matter transformations increase with watershed area, water residence time, and Damköhler numbers in nested watersheds” (v2)

This data package is associated with the publication “Riverine dissolved organic matter transformations increase with watershed area, water residence time, and Damköhler numbers in nested watersheds” submitted to Biogeochemistry by Ryan et al., 2024 (DOI: https://doi.org/10.1007/s10533-024-01169-5). This study aims to investigate fundamental and transferable drivers of dissolved organic matter (DOM) diversity across five nested watersheds within the contiguous United States. DOM diversity was explored using ultrahigh-resolution Fourier transform ion cyclotron resonance mass spectrometry (FTICR-MS). The samples and the unprocessed FTICR-MS data used in this study are publicly available on the Environmental System Science Data Infrastructure for a Virtual Ecosystem (ESS-DIVE) data repository (see DOIs below). The data for the Willamette, Gunnison, Connecticut, and Deschutes basins were collected as part of a collaboration between the Watershed Rules of Life (WROL) project and Worldwide Hydrobiogeochemistry Observation Network for Dynamic River Systems (WHONDRS). The data for the Yakima River basin (YRB) was collected by the PNNL River Corridor SFA. The raw, unprocessed FTICR-MS data with additional (meta)data can be found at doi:10.15485/1895159 for WROL samples and doi:10.15485/1898912 for YRB samples. This data package contains the processed data used in the associated manuscript. This package also contains ancillary geospatial, hydrological, and geochemical information that supports the interpretation of the FTICR-MS data within Ryan et al., 2024. This data package is associated with the GitHub repository found at https://github.com/WHONDRS-Hub/rcsfa-RC4-WROL-YRB_DOM_Diversity. This data package was originally published August 2024. It was updated January 2025 (modified files). See the change history in the readme more details. At the directory level, the data package is comprised of three folders: (1) data, (2) output, and (3) src; and five additional files including the data dictionary (file ending in "_dd.csv”) and file-level metadata (file ending in “_flmd.csv”). The “src” folder contains the scripts used to process the FTICR data, conduct the analyses, and produce the manuscript figures. The inputs for these scripts are in the “data” folder and the returned outputs in the “output” folder. Inputs include temporal and spatial metadata associated with the sampling efforts, processed FTICR data, and total and normalized putative biochemical transformations per sample. Outputs include cleaned and combined data presented as tables, descriptive statistics, and plots. The file-level metadata file lists all files contained in this data package and descriptions for each. The data dictionary describes the units and definitions for each tabular data column or row header.

54 ENVIRONMENTAL SCIENCES↗

Globally Gridded Groundwater Extraction Volumes and Costs under Six Depletion and Ponded Depth Targets

This repository contains simulated outputs from superwell – a hydro-economic tool for long-term assessment of groundwater cost and supply – providing globally gridded groundwater extractable volumes and associated unit costs ($/km³) for accessible groundwater production, based on a variety of user-defined depletion and ponded depth scenarios. Key model documentation: Niazi, H., Ferencz, S. B., Graham, N. T., Yoon, J., Wild, T. B., Hejazi, M., Watson, D. J., & Vernon, C. R. (2025). Long-term hydro-economic analysis tool for evaluating global groundwater cost and supply: Superwell v1.1. Geoscientific Model Development, 18(5), 1737-1767. https://doi.org/10.5194/gmd-18-1737-2025 Find the source code of the superwell model on GitHub: https://github.com/JGCRI/superwell Repository Overview Main output: superwell_outputs.7z contains 6 files (4.5 GB) named as superwell_py_deep_all_0.*PD_0.*DL.csv. These files present superwell outputs of global groundwater extraction volumes and cost estimates on a 0.5° scale for six scenarios with different Ponded Depth (PD; 0.3 and 0.6 m) and Depletion Limit (DL; 5%, 25%, and 40% of available volume) targets over the entire pumping lifetime of a grid cell superwell_py_deep_all_0.3PD_0.25DL_sample_100.csv contains superwell outputs for 100 data points sampled to match the global inputs' distribution superwell_py_deep_all_0.3PD_0.25DL_Grid_72548.csv contains superwell output for a single grid cell concept_v5.png provides an overview of the superwell workflow Outputs Description year_number: year of pumping depletion_limit: set depletion limit (DL) as a volume fraction of total available groundwater Mappings: continent, country, gcam_basin_id, Basin_long_name, grid_id: geographic identifiers and basin information Inputs: grid_area (km²): area of the grid cell whyclass: hydrogeological classification of the aquifer permeability (m/day), porosity (%), total_thickness (m), depth_to_water (m): aquifer properties. The geo-processed input data has been published separately: https://doi.org/10.57931/2307831 Model outputs: orig_aqfr_sat_thickness (m), aqfr_sat_thickness (m): original and remaining/instantaneous saturated thickness of the aquifer hydraulic_conductivity (m/day), transmissivity (m²/day): hydraulic properties of the aquifer radius_of_influence (m), areal_extent (km²): well radius and area of influence from the center of the well number_of_wells (-): number of wells in a grid cell determined by a ratio of well area and grid area max_drawdown (m), drawdown (m), drawdown_interference (m): well and aquifer drawdown during extraction total_head (m): total lift for the groundwater (depth to water plus drawdown) total_well_length (m): total depth of wells drilled well_yield (m³/day): pumping rate or well yield power (kW), energy (kWh): power and energy required for pumping groundwater Volume Outputs: volume_produced_perwell (m³), cumulative_vol_produced_perwell (m³): production volume metrics per well volume_produced_allwells (m³), cumulative_vol_produced_allwells (m³): aggregate extraction volumes for all wells in a grid cell available_volume (m³): available groundwater in storage for the grid cell as determined by aquifer properties depleted_vol_fraction: fraction of total volume pumped over available volumes in a grid cell (same as depletion limit) Cost Outputs: well_installation_cost ($): well installation cost based on the hydrogeological complexity of the aquifer annual_capital_cost, maintenance_cost, nonenergy_cost ($): nonenergy costs energy_cost_rate ($/kWh): electricity rate energy_cost ($): energy cost of pumping groundwater total_cost_perwell ($), total_cost_allwells ($): total annual energy and non-energy cost for each and all wells in a grid cell a unit_cost ($/m³), unit_cost_per_km3 ($/km³), unit_cost_per_acreft ($/acre-ft): total cost of pumping a unit of groundwater, indicated for different spatial units Key Resources Model documentation: Niazi, H., Ferencz, S., Graham, N., Yoon, J., Wild, T., Hejazi, M., Watson, D., & Vernon, C. (2024; In-prep). Long-term Hydro-economic Assessment Tool for Evaluating Global Groundwater Cost and Supply: Superwell v1. Geoscientific Model Development. Input data: Niazi, H., Watson, D., Hejazi, M., Yonkofski, C., Ferencz, S., Vernon, C., Graham, N., Wild, T., & Yoon, J. (2024). Global Geo-processed Data of Aquifer Properties by 0.5° Grid, Country and Water Basins. MSD-LIVE Data repository. https://doi.org/10.57931/2307831 superwell source code: https://github.com/JGCRI/superwell Cite as Niazi, H., Ferencz, S., Yoon, J., Graham, N., Wild, T., Hejazi, M., Watson, D., & Vernon, C. (2024). Globally Gridded Groundwater Extraction Volumes and Costs under Six Depletion and Ponded Depth Targets. MSD-LIVE Data repository. https://doi.org/10.57931/2307832 Contact Reach out to Hassan Niazi or Stephen Ferencz or open an issue in the superwell repository for questions or suggestions.

Earth Systems↗

The EGS Collab Project – Stimulations at Two Depths

The EGS Collab project, supported by the US Department of Energy, is performing intensively monitored rock stimulation and flow tests at the 10-m scale in an underground research laboratory to address challenges in implementing enhanced geothermal systems (EGS). Data and observations from the field tests are compared to simulations to understand processes and build confidence in numerical modeling of the processes. We have completed Experiment 1 (of 3), which examined hydraulic fracturing in a well-characterized underground fractured phyllite test bed at a depth of approximately 1.5 km at the Sanford Underground Research Facility (SURF) in Lead, South Dakota. Testbed characterization included fracture mapping, borehole acoustic and optical televiewers, full waveform sonic, conductivity, resistivity, temperature, campaign p- and s-wave investigations and electrical resistance tomography. Borehole geophysical techniques including passive seismic, continuous active source seismic monitoring, electrical resistance tomography, fiber-based distributed strain, distributed temperature, and distributed acoustic monitoring, were used to carefully monitor stimulation events and flow tests. More than a dozen stimulations and nearly one year of flow tests were performed. Quality data and detailed observations were collected and analyzed during stimulation and water flow tests using ambient temperature and chilled water. We achieved adaptive control of the tests using real-time monitoring and rapid dissemination of data and near-real-time simulation. More detailed numerical simulation was performed to answer key experimental design questions, forecast fracture propagation trajectories and extents, and analyze and evaluate results. Data are freely available from the Geothermal Data Repository. Experiment 2 examines the potential for hydraulic shearing in amphibolite at a depth of about 1.25 km at SURF. This site has a different set of stress and fracture conditions than Experiment 1. The Experiment 2 testbed consists of nine subhorizontal boreholes configured in two fans of two boreholes which surround the testbed and contain grouted-in electrical resistance tomography, seismic sensors, active seismic sources and distributed fiber sensors. A “five-spot” set of test wells that extends from a custom mined alcove includes an injection well and four production/monitoring wells. The testbed was characterized geophysically and hydrologically, and three stimulations have been performed using the Step-Rate Injection Method for Fracture In-Situ Properties (SIMFIP) tool to measure strains, and a new strain quantifying tool (downhole robotic strain analysis tool -DORSA) was deployed in a monitoring hole during stimulation. Real-time data were broadcast during stimulations to allow real-time response to arising issues.

EGS Collab, Enhanced Geothermal Systems, EGS, fiel↗

Observation of spin‑wave altermagnetic splitting in MnF2

-Contents of the Data Repository: - Polarized neutron diffraction data acquired in the (HK0) scattering plane. - Inelastic neutron scattering (INS) data from both unpolarized and polarized measurements with an incident neutron energy of Ei =9 meV.- - Reduced multidimensional single-crystal datasets (MDE) and corresponding S(Q,ω) slices used to generate all figures presented in the manuscript. - Julia source code and supporting input files used for spin-wave calculations, model fitting, and simulation of neutron scattering intensity maps.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Secure Data Logging and Processing with Blockchain and Machine Learning (Final Report)

Secure Data Logging and Processing with Blockchain and Machine Learning (ML) research is focused on the development of a platform to securely log and process sensor data in fossil power plants. The platform integrates two emerging technologies, blockchain and ML, and incorporates several innovative mechanisms to ensure the integrity, reliability, and resiliency of power systems. The goal is to protect the power plant from various cyberattacks such as false data injection and denial of service attacks using these technologies. The research goal was enabled by the following Research Project Objectives: 1) Secure authentication and identity verification of sensor nodes, actuators, and other equipment within a network. 2) Development of mechanisms that ensure only data sent by legitimate sensors are accepted and stored in the data repository. 3) Development of data aggregation methodologies using ML / Deep Learning (DL) algorithms to minimize noise / faulty data. 4) Implementation of the blockchain technologies to provide data security using secured IOTA framework & nodes.

20 FOSSIL-FUELED POWER PLANTS↗

Enabling pan-repository reanalysis for big data science of public metabolomics data

Public untargeted metabolomics data is a growing resource for metabolite and phenotype discovery; however, accessing and utilizing these data across repositories pose significant challenges. Therefore, here we develop pan-repository universal identifiers and harmonized cross-repository metadata. This ecosystem facilitates discovery by integrating diverse data sources from public repositories including MetaboLights, Metabolomics Workbench, and GNPS/MassIVE. Our approach simplified data handling and unlocks previously inaccessible reanalysis workflows, fostering unmatched research opportunities.

El Abiead, Yasin↗