Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “open data format”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Formation of Point Shocks for 3D Compressible Euler

We consider the 3D isentropic compressible Euler equations with the ideal gas law. We provide a constructive proof of the formation of the first point shock from smooth initial datum of finite energy, with no vacuum regions, with nontrivial vorticity present at the shock, and under no symmetry assumptions . We prove that for an open set of Sobolev‐class initial data that are a small L ∞ perturbation of a constant state, there exist smooth solutions to the Euler equations which form a generic stable shock in finite time. The blowup time and location can be explicitly computed, and solutions at the blowup time are smooth except for a single point , where they are of cusp‐type with Hölder C 1/3 regularity. Our proof is based on the use of modulated self‐similar variables that are used to enforce a number of constraints on the blowup profile, necessary to establish global existence and asymptotic stability in self‐similar variables. © 2022 Wiley Periodicals LLC.

Mathematics↗

MTUQ: a framework for estimating moment tensors, point forces, and their uncertainties

SUMMARY We introduce MTUQ, an open-source Python package for seismic source estimation and uncertainty quantification, emphasizing flexibility and operational scalability. MTUQ provides MPI-parallelized grid search and global optimization capabilities, compatibility with 1-D and 3-D Green’s function database formats, customizable data processing, C-accelerated waveform and first-motion polarity misfit functions, and utilities for plotting seismic waveforms and visualizing misfit and likelihood surfaces. Applicability to a range of full- and constrained-moment tensor, point force, and centroid inversion problems is possible via a documented application programming interface, accompanied by example scripts and integration tests. We demonstrate the software using three different types of seismic events: (1) a 2009 intraslab earthquake near Anchorage, Alaska; (2) an episode of the 2021 Barry Arm landslide in Alaska; and (3) the 2017 Democratic People’s Republic of Korea underground nuclear test. With these events, we illustrate the well-known complementary character of body waves, surface waves, and polarities for constraining source parameters. We also convey the distinct misfit patterns that arise from each individual data type, the importance of uncertainty quantification for detecting multimodal or otherwise poorly constrained solutions, and the software’s flexible, modular design.

58 GEOSCIENCES↗

Topography and canopy cover influence soil organic carbon composition and distribution across a forested hillslope in the discontinuous permafrost zone

This dataset contains data used for the paper "Topography and canopy cover influence soil organic carbon composition and distribution across a forested hillslope in the discontinuous permafrost zone". The Related References field will be updated with a full citation when available. Topography and canopy cover influence ground temperature in warming permafrost landscapes, yet soil temperature heterogeneity introduced by meso-topographic slope positions, microtopographic differences in vegetation cover, and the subsequent impact of contrasting temperature conditions to soil organic carbon (SOC) dynamics are understudied. Buffering of permafrost-affected soils against warming air temperatures in boreal forests can reflect surface soil characteristics (e.g., thickness of organic material) as well as the degree and type of canopy cover (e.g., open cover vs closed cover). Both landscape and soil properties interact to determine meso- and micro-scale heterogeneity of ground warming. We sampled a hillslope catena transect in a discontinuous permafrost zone near Fairbanks, Alaska to test the small-scale (1 to 3 meter) impacts of slope position and cover type on soil organic matter composition. Mineral active layer samples were collected from backslope, low backslope, and footslope positions at depths spanning 19 to 60 cm. We examined soil mineralogical composition, soil moisture, total carbon and nitrogen content, and organic mat thickness in conjunction with an assessment of SOC composition using Fourier-transform ion Cyclotron Resonance Mass Spectrometry (FT-ICR-MS). Soils in the footslope position had a higher relative contribution of lignin-like compounds while backslope soils had more aliphatic and condensed aromatic compounds as determined by FT-ICR-MS. The effect of open versus closed tree canopy cover varied with slope position. On the backslope, we found higher oxidation of molecules under open cover compared with closed cover, indicating an effect of warmer soil temperature on decomposition. Little to no effect of canopy was observed for soils at the footslope position, which we attributed, in part, to the strong impact of soil moisture content in SOC dynamics in the water-gathering footslope position. The thin organic mat under open cover on the backslope position may have contributed to differences in soil temperature and thus SOC oxidation under open and closed canopy. Here, the thinner organic mat did not appear to buffer the underlying soil against warm season air temperatures and thus increased SOC decomposition as indicated by higher oxidation of SOC molecules and a lower contribution of simple molecules under open cover compared with the closed canopy sites. Our findings suggest that the role of canopy cover in SOC dynamics varies as a function of landscape position and soil properties, namely organic mat thickness and soil moisture. Condition-specific heterogeneity of SOC composition under open and closed canopy cover highlights the protective effect of canopy cover for soils on backslope positions. This dataset contains a compressed (.zip) archive of the data and R scripts used for this manuscript. The dataset includes files in .csv format, which can be accessed and processed using MS Excel or R. This archive can also be accessed on GitHub at https://github.com/Erin-Rooney/Y1_fairbanks (DOI: 10.5281/zenodo.8071247).

54 ENVIRONMENTAL SCIENCES↗

Machine learning model inputs, outputs, and scripts associated with “Artificial intelligence-guided iterations between observations and modeling significantly improve environmental predictions”

NOTE: The manuscript associated with this data package is currently in review. The data may be revised based on reviewer feedback. Upon manuscript acceptance, this data package will be updated with the final dataset and additional metadata. This data package is associated with the manuscript “Artificial intelligence-guided iterations between observations and modeling significantly improve environmental predictions” (Malhotra et al., in prep). This effort was designed following ICON (integrated, coordinated, open, and networked) principles to facilitate a model-experiment (ModEx) iteration approach, leveraging crowdsourced sampling across the contiguous United States (CONUS). New machine learning models were created every month to guide sampling locations. Data from the resulting samples were used to test and rebuild the machine learning models for the next round of sampling guidance. Associated sediment and water geochemistry and in situ sensor data can be found at https://data.ess-dive.lbl.gov/datasets/doi:10.15485/1923689, https://data.ess-dive.lbl.gov/datasets/doi:10.15485/1729719, and https://data.ess-dive.lbl.gov/datasets/doi:10.15485/1603775. This data package is associated with two GitHub repositories found at https://github.com/parallelworks/dynamic-learning-rivers and https://github.com/WHONDRS-Hub/ICON-ModEx_Open_Manuscript. In addition to this readme, this data package also includes two file-level metadata (FLMD) files that describes each file and two data dictionaries (DD) that describe all column/row headers and variable definitions. This data package consists of two main folders (1) dynamic-learning-rivers and (2) ICON-ModEx_Open_Manuscript which contain snapshots of the associated GitHub repositories. The input data, output data, and machine learning models used to guide sampling locations are within dynamic-learning-rivers. The folder is organized into five top-level directories: (1) “input_data” holds the training data for the ML models; (2) “ml_models” holds machine learning (ML) models trained on the data in “input_data”; (3) “examples” contains files for direct experimentation with the machine learning model, including scripts for setting up “hindcast” run; (4) “scripts” contains data preprocessing and postprocessing scripts and intermediate results specific to this data set that bookend the ML workflow; and (5) “output_data” holds the overall results of the ML model on that branch. Each trained ML model resides on its own branch in the repository; this means that inputs and outputs can be different branch-to-branch. There is also one hidden directory “.github/workflows”. This hidden directory contains information for how to run the ML workflow as an end-to-end automated GitHub Action but it is not needed for reusing the ML models archived here. Please see the top-level README.md in the GitHub repository for more details on the automation. The scripts and data used to create figures in the manuscript are within ICON-ModEx_Open_Manuscript. The folder is organized into four folders which contain the scripts, data, and pdf for each figure. Within the “fig-model-score-evolution” folder, there is a folder called “intermediate_branch_data” which contains some intermediate files pulled from dynamic-learning-rivers and reorganized to easily integrate into the workflows. NOTE: THIS FOLDER INCLUDES THE FILES AT THE POINT OF PAPER SUBMISSION. IT WILL BE UPDATED ONCE THE PAPER IS ACCEPTED WITH ANY REVISIONS AND WILL INCLUDE A DD/FLMD AT THAT POINT. We thank the United States Forest Service, Washington Department of Fish and Wildlife, Washington Department of Natural Resources, Cowiche Canyon Conservatory, Washington State Parks and Recreation Commission (Scientific Research Permit #210901), and the Confederated Tribes and Bands of the Yakama Nation for access to field locations where the samples labeled “SSS” were collected. We also thank the Yakama Nation Tribal Council and Yakama Nation Fisheries for working with us to facilitate sample collection and optimization of data usage according to their values and worldview. WHONDRS consortium members were asked to provide any acknowledgments for the collection of samples labeled “CM” and the following is a list of acknowledgments that were submitted with their corresponding Site IDs: (MART) Research activities were conducted in part on the Wind River Experimental Forest within the Gifford Pinchot National Forest; (MP- 100379) Philadelphia is part of Lenapehoking, the ancestral homelands of the Lenape peoples; (MP-102398) Land surveyed is the ancestral homelands of the Nookhose'iinenno (Arapaho), Tsis tsis'tas (Cheyenne), and Nuuchu (Ute); (MP-100749 and MP- 100747) Georgia Coastal Ecosystem LTER, OCE-1832178; (SP-70 and SP-72) Eastern Shoshone, Shoshone-Bannock; (MP- 102944) Funded by Oregon Watershed Enhancement Board. On the traditional lands of the Confederated Tribes of the Siletz, Confederated Tribes of the Grand Rhonde, and the Clatsop-Nehalem Confederated Tribe; (MP- 100607) Holiday Creek is located on the traditional territory of the Monacan Indian Nation; (SP-45) Lafayette Blue Springs State Park; (MP-102420) NSF DEB-2016749; (MP-100019) New Hampshire Agriculture Experiment Station; (SP-35) Rayonier (land owner; https://www.rayonier.com/); (MP- 101276) US Department of Energy, Office of Science, Biological and Environmental Research, Subsurface Biogeochemical Research, Watershed Dynamics and Evolution SFA at ORNL; (MP- 103224) Watershed Dynamics and Evolution SFA at ORNL; (MP- 101584) Traditional lands of the Oceti Sakowin (Dakota, Lakota, Nakoda) and Anishinaabe Peoples.

54 ENVIRONMENTAL SCIENCES↗

Bond-centric modular design of protein assemblies

Directional interactions that generate regular coordination geometries are a powerful means of guiding molecular and colloidal self-assembly, but implementing such high-level interactions with proteins remains challenging due to their complex shapes and intricate interface properties. Here we describe a modular approach to protein nanomaterial design inspired by the rich chemical diversity that can be generated from the small number of atomic valencies. We design protein building blocks using deep learning-based generative tools, incorporating regular coordination geometries and tailorable bonding interactions that enable the assembly of diverse closed and open architectures guided by simple geometric principles. Experimental characterization confirms the successful formation of more than 20 multicomponent polyhedral protein cages, two-dimensional arrays and three-dimensional protein lattices, with a high (10%–50%) success rate and electron microscopy data closely matching the corresponding design models. Due to modularity, individual building blocks can assemble with different partners to generate distinct regular assemblies, resulting in an economy of parts and enabling the construction of reconfigurable networks for designer nanomaterials.

Biomaterials – proteins↗

Model Data Archive for Manuscript Titled "Evaluation of a Coupled Surface–Subsurface Hydrologic Model Using Dense Water‑Level Sensors in a Mixed Urban–Rural Watershed"

This archive provides scripts, input files, and datasets used for the implementation and evaluation of a fully coupled surface–subsurface hydrologic model in the Neches River Basin, southeast Texas. The study uses the Advanced Terrestrial Simulator (ATS) to simulate coupled surface–subsurface hydrologic processes over a mixed urban–rural watershed and evaluates model performance using a dense network of 136 in situ water-level sensors, nine U.S. Geological Survey (USGS) stream gauges, and SSEBop-derived evapotranspiration estimates during the period October 2014–June 2024. The workflow is implemented primarily in Python 3 using the Watershed Workflow package. The Jupyter notebooks can be executed using open-source software such as Anaconda JupyterLab or Visual Studio Code. Other data files include TXT, CSV, XML, SHP, TIF, NetCDF, HDF5, and ExodusII files, which can be processed using the provided Python scripts. ATS input files are provided in XML format and can be edited using any commonly used text editor. This archive contains: *Scripts and input files used to generate the ATS model setup, including watershed discretization, mesh generation, parameter mapping, and model configuration. *Jupyter notebooks used for preprocessing observational data, evaluating streamflow, water levels, and evapotranspiration, computing performance metrics, and generating the figures presented in the manuscript. *ATS simulation outputs and processed observational datasets, including OneRain and DD6 water-level sensors, USGS streamflow observations, GIS data, and supporting spatial datasets used throughout the study.

Dense water-level sensor network↗

Site and endmember spectra of terrestrial vegetation and soils for the Colorado Headwaters Ecological Spectroscopy Study, June-July 2025

This dataset provides site and endmember spectra collected during the 2025 Colorado Headwaters Ecological Spectroscopy Study (CHESS) campaign. The site spectra were collected to help validate airborne hyperspectral data acquired by the National Ecological Observatory Network's aerial observation platform (NEON AOP). Endmember spectra were collected to augment existing spectral libraries with additional samples of bare surfaces and non-photosynthetic vegetation. All measurements were acquired with an Analytical Spectral Devices (ASD) FieldSpec4 Hi-Res NG (Next Generation) spectroradiometer, which records radiance at 1nm (nanometer) intervals from the ultraviolet to the short-wave infrared (350-2500 nm). The dataset includes spectra measured at meadow sites where the CHESS team also collected vegetation samples for trait analyses. The site spectra were collected with the ASD FieldSpec4 palm grip attachment using an 8° field-of-view foreoptic. Site spectra are integrated measurements of the entire surface within the foreoptic’s field of view. For site-level spectra, the sun is the illumination source. A Spectralon panel mounted on a tripod was used for instrument optimization and white reference measurements for all site spectra. Site spectra were acquired within two hours of solar noon and within 48 hours of a NEON AOP overflight. Site spectra are labeled by date, sampling area, and site number according to the naming conventions of the CHESS campaign’s data management plan. The dataset also contains endmember spectra in the following categories: photosynthetic vegetation (PV), non-photosynthetic vegetation (NPV), bare (soil/rock), and flowers. Endmember measurements were acquired using either the contact probe or the leaf clip attachments of the ASD FieldSpec4. In these configurations, the bulb inside the spectrometer provides the light source for the measurements. The spectrometer was optimized and white reference measurements were recorded using the circular white pucks attached to the contact probe and leaf clip. Because they do not rely on solar illumination, contact probe and leaf clip measurements were collected during a broader time frame than the palm grip site spectra. Some endmembers were measured at CHESS meadow sites, while others were collected within the larger sampling area or in nearby locations (e.g. Gothic Townsite) with similar characteristics. Radiance, reflectance, and metadata files are split into three subfolders according to measurement type: proximal/palm grip (prx), contact probe (cp), and leaf clip (lc). Radiance spectra are provided in ASD file format (.asd file extension). All ASD files can be opened using the provided scripts. Metadata is provided in two formats: CSV file format (no geolocation) and GEOJSON file format (includes geolocation for each spectra). The dataset includes a set of pre-processed reflectance spectra as CSV files (yyyymmdd_rfl.csv). The python scripts and jupyter notebook used to calculate reflectance spectra from the ASD radiance data is included here and was previously published at: https://doi.org/10.3334/ORNLDAAC/2446. There is also a folder of JPEG photographs corresponding to selected spectra. We include a protocol document with detailed steps for ASD FieldSpec4 assembly and operations. This data additionally contains a file level metadata (flmd.csv) and data dictionary (dd.csv) file. Geospatial information: Geospatial data for mapping measurement site locations are in the files CHESS_polygons_lai_UTM.geojson, CHESS_polygons_shrub_UTM.geojson, and CHESS_polygons_meadow_UTM.geojson in the companion geospatial package for the 2025 CHESS campaign, ‘CHESS 2025: Location data for field observations and sampling’ (Henderson et al., 2026). CHESS Project Description: The Colorado Headwaters Ecological Spectroscopy Study (CHESS) comprised a multi-week airborne remote sensing and field observation campaign in the Upper Gunnison Basin, Colorado, conducted in June and July of 2025. Airborne remote sensing was conducted by the National Ecological Observatory Network Airborne Observation Platform (NEON AOP), concurrent with a field campaign run by the Rocky Mountain Biological Laboratory (RMBL), the Lawrence Berkeley National Laboratory (LBNL) and SLAC National Accelerator Laboratory Watershed Function Science Focus Area (SFA), and NASA-JPL (Jet Propulsion Laboratory) Earth Surface Mineral Dust Source Investigation (EMIT) program. Between June 10 and July 18, 2025, the NEON AOP flight team collected high-resolution aerial imaging spectroscopy and Light Detection and Ranging (LiDAR) data over three domains: the Upper East River (CRBU), Almont Triangle (ALMO), and the Upper Taylor Basin (UPTA). In coordination with the flights, a field campaign acquired ground-truth observations, including observations of vegetation composition, foliar traits, forest demography, and subsurface properties in 18 core sampling areas within the domains. Additional surface water observations were taken at over 380 point locations. All CHESS campaign datasets can be found within the CHESS ESS-DIVE data portal: https://data.ess-dive.lbl.gov/portals/chess. Funding Acknowledgment: This research was carried out at the Jet Propulsion Laboratory, California Institute of Technology, under a contract with the National Aeronautics and Space Administration (80NM0018D0004) and was funded by EMIT Extended Mission Phase E Science.

2018 NEON and 2025 CHESS Campaigns↗

Computational Tools and Workflows for Quantitative Risk Assessment and Decision Support for Geologic Carbon Storage Sites: Progress and Insights from the U.S. DOE’s National Risk Assessment Partnership

The 2005 Intergovernmental Panel on Climate Change (IPCC) Special Report on CCS raised the profile of CO2 capture and storage (CCS) as an important technology for reducing greenhouse gas (GHG) emissions. CCS is now recognized as a key component of most climate change mitigation scenarios. Since publication of that report the international research, development, and deployment (RD&D) community has advanced key technical aspects, clarified regulatory requirements, explored value chain and infrastructure solutions, and developed incentive paradigms to enable and promote large-scale deployment of CCS. These efforts have included research to better characterize geologic storage resources, to improve injection performance and storage efficiency, to assess and manage subsurface environmental risks, and to advance monitoring technologies to assure system conformance. These efforts have helped to build confidence in the viability of geologic carbon storage (GCS), but stakeholder concerns about long-term risks and liability associated with GCS remain a hurdle to broad acceptance and large-scale deployment of CCS. Since 2010, the U.S. DOE’s National Risk Assessment Partnership (NRAP) – a research collaboration between five contributing national laboratories – has worked to establish and demonstrate methods and tools to quantify and manage the subsurface environmental risks associated with GCS, amidst uncertainty. This work supports the Office of Fossil Energy and Carbon Management Carbon Transport and Storage Program’s goal of advancing safe and secure commercial-scale GCS deployment. To address the technical challenge of simulating the physical response of the GCS site to large-scale CO2 injection, NRAP has adopted an approach that relies on coupling computationally efficient reduced-order and/or data-driven proxy models of important system components (i.e., storage reservoir, sealing caprock, leakage pathways, intermediate formations, overlying groundwater aquifers, and the atmosphere) in integrated assessment framework. That integrated model of the physical system is complemented with fit-for purpose functionality to support site characterization and risk-related decisions. The recently released NRAP Phase II toolset includes the Open-Source Integrated Assessment Model (NRAP-Open-IAM) for evaluation of trends in leakage risk and potential impact, tools to support monitoring design optimization (Designs for Risk Evaluation and Management – DREAM v3.0 and Passive Seismic Monitoring Tool - PSMT), and tools for state of stress evaluation (State-of-Stress Analysis Tool - SOSAT) and forecasting induced seismicity risk. The NRAP team has also released a pair of reports describing conceptual workflows to incorporate physics-based, quantitative risk assessment into many of the design, planning, operation, and closure decisions for GCS projects. An online catalogue highlights published studies where these tools and methods are demonstrated. In this presentation, the utility of these products to assess risks and address key stakeholder questions will be highlighted through examples, and related insights about the safety and security of geologic carbon storage in qualified storage sites will be discussed. The prospect of rapid, large-scale deployment of GCS technology to aggressively reduce anthropogenic CO2 emissions requires careful consideration of interference between multiple commercial-scale storage projects within a basin. Going forward, NRAP is expanding and adapting site-scale risk quantification tools and methods to enable assessment of risks and inform management decisions for basin-scale deployment. Increasingly, this work will leverage next-generation approaches for surrogate modelling, fast prediction, and advanced visualization enabled by machine learning and artificial intelligence to promote virtual learning, scenario evaluation, and augment risk-based decision making.

quantitative risk assessment, geologic carbon stor↗

Keyhole-mode Microscopy Dataset for Laser Powder-bed Fusion Modeling

Laser powder-bed fusion (LPBF) is an additive manufacturing (AM) technology that uses high-power sintering lasers to precisely construct metal designs. Material is accumulated by selectively sintering regions of a metal powder layer to a growing structure underneath forming a 3D geometry. Certain conditions make fusion in this process to operate in "keyhole-mode," characterized by targeted materials evaporating as plasma. While keyhole-mode operations can produce deeper molten pools than that observed in "conduction-mode", this mode of operation is often undesirable, as its molten regions can collapse on themselves, encapsulating metal vapors and forming cavities. The formation of such cavities negatively affects strength and consistency of the fused materials. To identify and avoid these detrimental effects of keyhole-mode operation, considerable data collection and analysis regarding this mode of operation is needed. Therefore, under the support of the "Open Data Initiative," Lawrence Livermore National Lab is releasing a dataset for analysis of this keyhole effect, and the conditions which transition operation from conduction-mode to keyhole-mode melting. This dataset consists of 600+ micrographs from laser powder-bed single track runs, with differing cross-section angles and input parameters, operating under both normal and keyhole mode operations. This journal documents the collection, organization, and usage of this dataset.

36 MATERIALS SCIENCE↗

Data from: 'Abiotic influences on continuous conifer forest structure across a subalpine watershed'

This package archives the core data used for analysis and inference in 'Abiotic influences on continuous conifer forest structure across a subalpine watershed' (Worsham et al., 2025). All data were collected in the East River, Washington Gulch, Slate River, and Coal Creek watersheds of Colorado. In the paper, we quantified the relative influence of climate, topographic, edaphic, and geologic factors on conifer stand structure and composition, and their functional relationships, at the watershed scale. We used waveform LiDAR data to derive spatially continuous stand structure metrics. We fused these with a species-level classification map to estimate tree species abundance. We applied generalized additive and generalized boosted models to evaluate the covariability of structural and compositional metrics with abiotic variables. The package contains the essential products required for reproducing our analysis and the tables and figures reported in the publication. The products comprise four classes: (1) geospatial data, (2) tabular data used for inferential analysis, (3) tabular data describing analytical results and performance statistics, and (4) a data user guide. (1) includes discretized waveform LiDAR data, locations and attributes of individual tree crowns, sampling locations and domain boundaries, a canopy height model, and raster files of estimated forest structural and compositional metrics at 100 m grid scale. (2) includes all response and explanatory variable values applied in inferential models. Response variables include conifer forest stand density, basal area, 95th percentile height, quadratic mean diameter, and others. Explanatory variables include climatic water deficit, actual evapotranspiration, elevation, heat load, soil available water content, and others. (3) includes results of training and testing several individual tree detection (ITD) algorithms, as well as inferential modeling results. (4) is a PDF user guide for this data package, including detailed descriptions and data dictionaries for all files. The data package root contains 17 assets: 8 compressed tape archive (.tar.gz) files, 5 comma-separated values (.csv) files, 3 Geographic Tagged Image File Format (GeoTIFF) (.tif) files, and 1 Portable Document Format (.pdf) file. The compressed .tar.gz archives contain ESRI shapefiles (.shp) .tif, compressed LASer (.laz), and .csv files. The archives must first be decompressed using the widely distributed command-line software utility TAR. All other files, including constituent files within the .tar.gz archives, can be opened in the open-source R statistical computing environment. Alternatively, .csv files may also be read in any simple text editor software or Microsoft Excel. Geospatial files including .shp and .tif files can also be opened in GIS software, such as QGIS (open-source) or ESRI ArcGIS (proprietary). The .pdf Data User Guide can be read with Adobe Acrobat Reader or other compatible readers.

2018 NEON and 2025 CHESS Campaigns↗

Automatic Calibration of a Geomechanical Model from Sparse Data for Estimating Stress in Deep Geological Formations

Summary In this study, we demonstrate geomechanical modeling with fully automatic parameter calibration to estimate the full geomechanical stress fields of a prospective US carbon dioxide (CO2) storage site, based on sparse measurement data. The goal is to compute full stress tensor field estimates (principal stresses and orientations) that are maximally compatible with observations within the constraints of the model assumptions, thereby extending pointwise, incomplete partial stress measurement to a simulated full formation stress field, as well as a rough assessment of the associated error. We use the Perch site, located in Otsego County, Michigan, USA, as our case study. The input data consist of partial stress tensor information inferred from in-situ borehole tests, geophysical well logs, and processing of seismic data. A static earth model (SEM) of the site was developed, and geomechanical simulation functionality of the open-source MATLAB Reservoir Simulation Toolbox (MRST) was used to model the stress field. Adjoint-based nonlinear optimization was used to adjust boundary conditions and material properties to calibrate simulated results of observations. Results were interpreted through a Bayesian framework. The focus of this paper is to demonstrate how the fully automatic calibration procedure works and discuss the results obtained; it does not attempt a detailed analysis of the stress field in the context of the proposed CO2 storage initiatives. Our work is part of a larger effort to noninvasively determine in-situ stresses in deep formations considered for CO2 storage. Guided by previously published research on geomechanical model calibration, our work presents a novel calibration approach supporting a potentially large number of linear or nonlinear calibration parameters to produce results optimally agreeing with available measurements and thus extend partial pointwise estimates to full tensor fields compatible with the physics of the site.

Engineering↗

How Does Land Cover and Its Heterogeneity Length Scales Affect the Formation of Summertime Shallow Cumulus Clouds in Observations From the US Southern Great Plains?

This study investigates the effects of heterogeneous land covers on shallow cumulus (ShCu) clouds at the US Southern Great Plains using high-resolution satellite and land cover data. During late summer, ShCu occurs over cities the most frequently and over open waters the least frequently, and more often over forest than over grassland. The preferential occurrence of ShCu over forest relative to grassland is consistent with surface measurements showing larger heat fluxes over forests. This preferential occurrence also varies with the length scales of land patches with the largest cloud occurrence difference shifting from smaller length scales (<9 km) during midday to larger scales (>9 km) in the early afternoon. Consistent with theory, these signals are more pronounced under low wind conditions. The preferential length scale shift with time suggests the existence of secondary circulations that strengthen and promote convergence over larger spatial scales as the differential land surface heating intensifies.

54 ENVIRONMENTAL SCIENCES↗

Alpha-decay width of a near proton-threshold resonance in the 7 Li( α, α ) channel

We investigate the alpha-decay width of a near proton-threshold resonance in 11 B by the excitation functions of 7 Li(α, α) and 7 Li(α, α’) reactions. This resonance is an example of loosely bound atomic nuclei, understood as small open quantum systems, that exhibit properties significantly influenced by their coupling to the continuum. The present experiment focuses on the formation of a specific state in 11 B at excitation energy in the region near 11.4 MeV by bombarding a 7 Li target with a 4 He beam at energies ranging from 3.92 to 4.56 MeV in the laboratory frame. The study provides data complementary to previous observations of the proton emission in the β-decay of the neutron-rich halo nucleus 11 Be. An R-matrix fit to the data provides resonance energy and partial widths consistent with J π =1/2 + for a narrow near-threshold state in 11 B at 11.400(20) MeV, that is 171(20) keV above the proton emission threshold, with partial α-decay width Γ α =2.5$^{+3.5}_{–1.5}$ keV in elastic scattering and Γ α' =15.8$^{+1.9}_{–0.4}$ keV in inelastic scattering. These findings contribute to the understanding of the structure of this open quantum system resonance.

11B↗

Improving the Accessibility and Usability of Geothermal Information with Data Lakes and Data Pipelines on the Geothermal Data Repository: Preprint

The Geothermal Data Repository (GDR) provides universal access to data and information resulting from research and development activities funded by the Department of Energy (DOE). The GDR has extended this universal access to big data through integration with data lakes developed by the Open Energy Data Initiative (OEDI). Previously, large datasets such as seismic waveform or distributed acoustic sensing (DAS) data could only be accessed by institutions with high performance data storage and compute capabilities, effectively limiting the accessibility of big data to national labs, larger universities, and major corporations. Moreover, the time and resources needed to transport big data and configure them can produce additional barriers to use. Many of the standard formats used for structured data models (also known as content models) are incapable of handling big data and can introduce additional usability problems, often requiring data to be reformatted prior to use. This paper will explore how recent integrations between the GDR and the OEDI data lake have improved the accessibility and usability of geothermal data in a big way, making the data available to a broader audience, and enabling collaborative analysis and innovation across the greater geothermal industry.

access↗

Improving the Accessibility and Usability of Geothermal Information with Data Lakes and Data Pipelines on the Geothermal Data Repository

The Geothermal Data Repository (GDR) provides universal access to data and information resulting from research and development activities funded by the Department of Energy (DOE). The GDR has extended this universal access to big data through integration with data lakes developed by the Open Energy Data Initiative (OEDI). Previously, large datasets such as seismic waveform or distributed acoustic sensing (DAS) data could only be accessed by institutions with high performance data storage and compute capabilities, effectively limiting the accessibility of big data to national labs, larger universities, and major corporations. Moreover, the time and resources needed to transport big data and configure them can produce additional barriers to use. Many of the standard formats used for structured data models (also known as content models) are incapable of handling big data and can introduce additional usability problems, often requiring data to be reformatted prior to use. This paper will explore how recent integrations between the GDR and the OEDI data lake have improved the accessibility and usability of geothermal data in a big way, making the data available to a broader audience, and enabling collaborative analysis and innovation across the greater geothermal industry.

access↗

PlasmoData.jl — A Julia framework for modeling and analyzing complex data as graphs

Datasets encountered in scientific and engineering applications appear in complex formats (e.g., images, multivariate time series, molecules, video, text strings, networks). Graph theory provides a unifying framework to model such datasets and enables the use of powerful tools that can help analyze, visualize, and extract value from data. In this work, we present PlasmoData.jl, an open-source, Julia framework that uses concepts of graph theory to facilitate the modeling and analysis of complex datasets. The core of our framework is a general data modeling abstraction, which we call a DataGraph. We show how the abstraction and software implementation can be used to represent diverse data objects as graphs and to enable the use of tools from topology, graph theory, and machine learning (e.g., graph neural networks) to conduct a variety of tasks. We illustrate the versatility of the framework by using real datasets: (i) an image classification problem using topological data analysis to extract features from the graph model to train machine learning models; (ii) a disease outbreak problem where we model multivariate time series as graphs to detect abnormal events; and (iii) a technology pathway analysis problem where we highlight how we can use graphs to navigate connectivity. Further, our discussion also highlights how PlasmoData.jl leverages native Julia capabilities to enable compact syntax, scalable computations, and interfaces with diverse packages. Overall, we show that the DataGraph abstraction and PlasmoData.jl Julia package are able to model data within graphs and enable useful analysis.

96 KNOWLEDGE MANAGEMENT AND PRESERVATION↗

MBARI-WEC September and October 2022 Field Data

This data is needed to simulate a model of the MBARI-WEC (Monterey Bay Aquarium Research Institute, Wave Energy Converter device) in a simulation environment (e.g. Gazebo) for 56 observation dates in the time between September and October 2022, and to compare the simulation outputs to the corresponding field data of the physical MBARI-WEC. There were 50 observations chosen in Sept and 6 observations in Oct. To help understand terms below, a summary of the system can be found at the github link in the downloads section below. The Gazebo MBARI-WEC model is also provided, should users wish to simulate using this platform. There are 4 *mat files included. ................................................................................................................................................................................................................................... Spectrum and Simulation Inputs: September2022_spectrum_siminputs.mat and October2022_spectrum_siminputs.mat has data needed for simulation inputs in table format. These include the ocean spectrum for an observation and operating parameters of the MBARI-WEC during that observation. They are organized as rows representing an observation and columns representing data. For example, for the September *mat there are 50 rows. The first 7 columns are Datetime, sig_waveheight, peak_period, mean_period, heaveconedoor_status, pistonpos_mean, and scale_factor: - Datetime is the date and time the observation occurred in PST - sig_waveheight is the significant wave height of the ocean spectrum during that observation in meters - peak_period is the peak period of the ocean spectrum during that observation in seconds - mean_period is the mean period of the ocean spectrum during that observation in seconds - heaveconedoor_status is the status of the heave cone doors where 0 represents the doors are open and 1 represents they are closed - pistonpos_mean is the mean position of the PTO ram (piston) in meters - scale_factor is an additional factor of 0.5 --1.4 applied to a default damping relationship The next columns are data needed to represent the ocean spectrum. First are the frequencies [Hz] labeled as "f0-f38", then the variance density [m2/Hz] labeled as "vardens0-vardens38". October2022_spectrum_siminputs.mat follows as a similar format as above, but includes a larger amount of ocean spectrum frequencies and variance density elements. ................................................................................................................................................................................................................................... Field data: The field data is found in MBARIWEC_septdata.mat and MBARIWEC_octdata.mat for the observations of September and October, respectively. These contain data in a struct format. The struct contains the following fields for each observation: PC_BattCurr, PC_LoadCurr, PC_RPM, PC_Voltage, SC_Range, SC_Velocity, DateTime, where: - PC_BattCurr is the current flowing to or from the onboard batteries in Amps - PC_LoadCurr is the current flowing to the load dump in Amps - PC_RPM is the electric/hydraulic motor shaft speed (directly coupled) in RPM - PC_Voltage is the bus voltage at the power converter in Volts - SC_Range is the PTO ram (piston) position in meters where 0 is fully retracted and 2.03 is fully extended - SC_Velocity is the PTO ram (piston) velocity in meters/sec - DateTime is the date and time of the sampled field data in each observation in PST - Electric Power is equal to: PC_Voltage*(PC_BattCurr + PC_LoadCurr) in Watts For example, upon loading MBARIWEC_octdata.mat, the aforementioned fields would be loaded, each with {6x1} cells for the 6 observations chosen in October. Within the first cell of e.g. SC_Range would be sampled data representing the field data of the MBARIWEC PTO piston position for say, one hour, of the first October observation. The corresponding field DateTime would...

16 TIDAL AND WAVE POWER↗

The "PVLib" of Degradation: PVDeg

The Photovoltaic (PV) industry constantly aims for lower costs through higher-efficiency cells, improved module designs, and improvements in durability. This leads to the use of new materials, designs, and manufacturing processes, and not always with a sufficient amount of durability testing. To help drive down costs there is a desire to create modules that will last for up to 50 years of service life. To accomplish this, every degradation mode and mechanism must be identified and either eliminated or otherwise mitigated. This involves the extrapolation of laboratory results to the field conditions. There is a need to organize the existing degradation data into an accessible format and to provide industry relevant tools for extrapolation from laboratory to field conditions. While the basic equations used to model degradation are sometimes very simple, the full analysis involves calculations are cumbersome but ubiquitous for many degradation processes. A simplified, modeling framework to accomplish these repetitive processes will facilitate the analysis to help researchers keep up with the rapid pace of technological changes. In this talk, we will describe our progress creating the open-source tool PVDeg. This tool can be used to search for and analyze degradation information and extrapolate PV module performance and durability to field exposure. PVDeg simplifies many of the common foundational computational operations for obtaining meteorological data and using it to generate a model of the PV deployment. This prediction tool repository also contains various degradation models as well as a library of material parameters suitable for estimating the durability assessment of materials and components. We use an integration pipeline approach that allows us to leverage weather data from the National Solar Radiation Database, and other weather sources, to perform geospatial degradation analysis in the US and worldwide. We hope to become a repository that can be used for weathering and degradation analysis for various applications beyond the PV industry. During the talk, we will provide the PVPMC attendees the opportunity to interact with the tool via a Google Collab tutorial they can run on their phones or laptops.

durability↗