Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “logging”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 145 records · Page 8

Quantifying Time to Charge

This work presents analysis aimed at quantifying and improving the time between the start of an electric vehicle (EV) charging session and the start of actual energy transfer. This is achieved by analyzing EV charge session communication logs. The data from these logs are used to identify the longest-duration phases within charge initialization. Paths for future work are presented with a focus on the areas most likely to deliver overall time-to-charge improvements.

33 ADVANCED PROPULSION SYSTEMS↗

Expedition UT-GOM2-2 Site H

Pressure and conventional cores were collected at Site H of the Walker Ridge Protracted Area Block 313 in the Terrebonne Basin, deepwater Gulf of America (Gulf of Mexico) during the University of Texas (UT) Deepwater Hydrate Coring Expedition (UT-GOM2-2). Pressure and conventional cores were collected continuously to a depth of 155.1 meters below the seafloor (mbsf). At deeper depths, cores were taken periodically from hydrate-bearing sands and their bounding muds to a total depth of 861.3 mbsf. 162.6 m of conventional core and 54.8 m of pressure core were recovered. Twelve temperature measurements were made between 27.1 and 144.5 mbsf to determine the geothermal gradient. At the seafloor, more than 4 m of sandy silt of unknown origin was encountered. Beneath this sand, to a depth of about 200 mbsf, the section was composed of interbedded mud and biogenic carbonate ooze. The ooze correlated to low density and high porosity intervals observed in the previously acquired logging while drilling (LWD) data and as measured. These ooze intervals also correspond to lighter sediment color, increased Ca content based on X-ray florescence (XRF) core scanning, and increased calcareous nannofossil abundance. Calcareous nannofossil biostratigraphy constrains the entire record to the Pleistocene (< 0.91 million years), with a pronounced increase in sedimentation rate with depth. Below 200 mbsf, the section was predominantly composed of mud with two thicker, hydrate-bearing coarse-grained intervals, which are commonly known as the Blue and Orange sands. The dissolved gas concentration was quantified from pressure cores. In the shallow section, dissolved methane concentration increased below the sulfate-methane transition zone (SMTZ) and reaches saturation (the limit of solubility for methane) at 147 mbsf. Gas expansion was very common in conventional and depressurized pressure (conventionalized) cores below the SMTZ. At deeper depths, the methane concentration within muds bounding the Blue and Orange reservoirs was generally found to be less than saturation. The dissolved and hydrate gas composition is consistent with a microbial source, containing greater than 99.99% methane and only trace concentrations of ethane, propane, and butane. The methane to ethane ratio (C1 /C2 ) and the methane to ethane plus propane (C1 /(C2 +C3 )) decrease with depth down to at least 678 mbsf, mainly driven by the increase in ethane with depth. It is unclear if this trend continues through the Orange sand interval. The δ13C isotopic signature of methane ranges between -69.9 and -78.5 ‰ relative to the Vienna Pee Dee Belemnite (VPDB) standard. Pressure core recovery in sandy intervals was poor. However, pressure core logs of the Orange sand show intervals of low density and high velocity, which are indicative of high hydrate saturation. One pressure core was degassed and the average hydrate saturation in the core was determined to be 24%. One core from within the Orange sand was composed of interbedded graded sandy silt and mud. The sandy silts from this core are composed of mainly quartz and feldspar with some lithics. Most of the recovered pressure core samples are maintained at near in-situ pressure and temperature (within the hydrate stability field) at the University of Texas Pressure Core Center awaiting analysis. In the shallow section, samples will be used to determine the flux of organic carbon through the basin system, find the rate at which that carbon was consumed, and understand the microbial population responsible for these processes. In the deeper section, samples from in and around the hydrate reservoirs will be used to determine the petrophysical properties of the reservoir and bounding seals in these systems.

03 NATURAL GAS↗

lllinois Storage Corridor CarbonSAFE Phase III: Pre-drilling Site Assessment: Prairie State Generating Company

The Illinois Storage Corridor project will drill a stratigraphic test well as part of the Illinois Storage Corridor CarbonSAFE Phase 3 project near the Prairie State Generating Company coal-fired power plant near Marissa, Illinois. The pre-drilling site evaluation has considered the primary target reservoirs, the Potosi Dolomite and St. Peter Sandstone, and primary seal, the Maquoketa Group. Data to be collected from the well include core, fluid samples, in situ well tests, geophysical logs intended to provide information on lithologic, geomechanical, and geophysical characteristics to determine the feasibility for the geologic sequestration of 50 million metric tons or more of injected carbon dioxide. The planned drilling site has been evaluated using available subsurface geologic data and analyses from the Illinois Basin. These data provide lithologic and structural information, shallow groundwater resource distribution, location of known nearby wellbores, and regional drilling characteristics. The data were used to generate geologic structure and isopach maps for the target reservoir and caprock strata and for prognosing the tops of major lithologic units to aid drilling and coring procedures. The regional analyses indicate that no known structural features are expected to negatively impact the target storage reservoir or caprock. No protected and sensitive areas, groundwater resources, or existing resource development are expected to be impacted by the proposed well drilling activities. The well is planned to be drilled to a total depth of approximately 5,600 feet (1,707 m) and terminate in the Precambrian. Cores (up to 5 intervals) will be collected from the Maquoketa Group, confining units above the St. Peter Sandstone, St. Peter Sandstone, confining units of the Potosi Dolomite and the Potosi Dolomite. Water samples will be attempted to be collected from the St. Peter Sandstone and Potosi Dolomite. Potential impact on drilling progress is a lost circulation zone in the Potosi Dolomite, which has been demonstrated to have intermittent cavernous porosity from karstification elsewhere in the Illinois Basin. This document also presents a preliminary coring and sampling program, proposed logging suite, and well testing program, all of which will be reviewed during drilling.

01 COAL, LIGNITE, AND PEAT↗

Roughrider Carbon Storage Hub (Final Report)

The Roughrider Carbon Storage Hub was a 2-year project (October 2023 – September 2025) conducted by the Energy & Environmental Research Center (EERC) focused on advancing the feasibility of a commercial-scale carbon dioxide (CO 2 ) geologic storage hub in McKenzie County, North Dakota. The project’s objective was to investigate the potential that stacked storage complexes (multiple deep saline formations) can safely and economically store at least 50 million tonnes of CO 2 within 30 years. The captured CO 2 would be sourced from industrial emitters including project partner ONEOK, Inc.’s gas-processing plants and a planned gas-to-liquids facility. Drilling of the Roughrider 1 stratigraphic test well (14,979-ft total depth) was completed in November 2024. The wellbore intersected four candidate storage formations: Inyan Kara, Broom Creek, Mission Canyon, and Black Island–Deadwood. Operational challenges, including a stuck drill string, were resolved without long-term impact. A comprehensive logging and coring program was conducted, followed by successful well abandonment and site reclamation. Over 660 ft of 4-in. whole core was retrieved. Core plug samples were processed and analyzed for petrophysical and geochemical properties. Results confirmed promising porosity and permeability in the Inyan Kara and Broom Creek Formations and removal of the Mission Canyon and Black Island–Deadwood horizons from further investigation. Data derived from the logging and coring program were used to improve initial geologic models built from legacy data. CO 2 injection simulations showed that the Inyan Kara alone can feasibly store the target mass of CO 2 . Because of subtle differences in geologic structure and porosity trends between the formations, a stacked storage scenario using the Broom Creek and Inyan Kara Formations resulted in a larger overall plume area than using the Inyan Kara alone. Preliminary CO 2 pipeline routes from the industrial sources were mapped utilizing existing rights of way and evaluated for capacity and cost using U.S. Department of Energy Office of Fossil Energy and Carbon Management/National Energy Technology Laboratory models and U.S. Environmental Protection Agency emissions data. Integrating capture, transport, and storage cost estimates with policy incentives (e.g., 45Q credits) provided a total cost-per-ton analysis. Results indicate that the small scale of the volumes to be transported over the cumulative large distances does not support the project’s financial viability. However, the groundwork laid during this project from geological, regulatory, and social perspectives positions the Roughrider hub site as a promising candidate for commercial carbon storage in North Dakota, especially if the economy of scale is introduced for CO 2 transportation to the hub site.

01 COAL, LIGNITE, AND PEAT↗

Structural Evolution of the Hogback Monocline and Its Tectonic Significance in the San Juan Basin

The San Juan Basin is recognized as a Laramide foreland basin. It is located within the Colorado Plateau, a broad tectonic province characterized by a thick sedimentary sequence that was segmented into smaller sub basins during the Late Cretaceous to Paleogene Laramide orogeny. The Hogback Monocline lies along the northwestern margin of the San Juan Basin and is considered a Laramide-age structure formed in response to compressional stress. In this study, we interpret surface and subsurface datasets to construct a structural geological model and evaluate its tectonic significance. Through seismic data, we identify key fault and fold geometries at depth. The seismic dataset used in this study was reprocessed in depth and constrained with well log velocity data to enhance seismic imaging quality. Additionally, we performed well log correlations to identify formation tops and assess variations in basin infill and thickness geometry. A series of structural cross-sections, constructed using seismic data and a high density of boreholes, are presented to evaluate geometric variations along the structure and its evolution during basin development. Furthermore, kinematic restoration and forward modeling analyses were conducted to validate our structural interpretation. This work suggests that the Hogback Monocline formed through fault-propagation folding and flexural slip affecting the pre-Laramide sedimentary sequence under compressional stresses associated with the Laramide orogeny. This structure is interpreted as a high-angle reverse fault that influenced the geometry of the late basin infill. Additionally, monocline bending along the structure may have been controlled by fault relay systems and, in some cases, influenced by strike-slip faulting.

Reyes, Martin [New Mexico Bureau o fGeology and Mi↗

Predicting Li-Ion Battery Capacity Fade Using Early-Life Data and a Hybrid Data-Driven Gaussian Process-Bayesian Regression Approach

Accurately predicting Li-ion battery capacity trajectories using early-life data can dramatically improve battery-life understandings and be used to rapidly evaluate design/cost/performance trade-offs when developing new battery materials. Accurate early-life predictions enable researchers to quickly iterate over cell designs and material precursor properties without consistently cycling cells to failure. To this end, we present a toolbox that uses a combined Gaussian Process and Bayesian regression approach that capitalizes on signals other than just capacity (e.g., dQ/dV, voltage drops) to rapidly predict capacity-fade trajectories. The prediction tool uses Bayesian regression to fit functional forms, e.g., power law, sigmoids, etc., to predict capacity-fade dynamics. By fitting functional forms, the capacity fade can be interrogated at any point in the future, allowing for early cell-failure prediction. Additionally, Bayesian regression allows for accurate uncertainty estimates that account for cell-to-cell variability (aleatoric uncertainty) and the lack of observation data (epistemic uncertainty). By only using early cycle data to predict the capacity fade trajectory, uncertainty bounds at end-of-life can be extremely large. The large uncertainty bounds are further exacerbated because there is no systematic way to define the prior distribution of the functional forms' parameters. We improve our the predicted trajectory confidence interval of our predicted trajectory using two methods. First, we shows that a small amount of held-out cycling data is sufficientuse some train cells, that have been cycled to failure to derive information regarding the appropriate prior distributions for the functional forms' parameters of the functional form, effectively leading to data-driven priors.. We propose constructing the data-driven priors by first running a Bayesian regression starting with uninformed priors to generate intermediate cell-specific posterior parameter distributions. These posterior distributions are combined using a Ggaussian mixture model for each parameter to create the data-driven priors. These mixture models serve as the data-driven prior distributions for the parameters for. Second, we derive multiple features, e.g., C_dchg 0.5 DoD 0.5, log (|mean(dQ/dV_(w_3-w_0 ) (V)|), etc., from the train cellsheld-out cycling data, identify which the features are that best predicting capacity at early/mid-life cycles, and then create Ggaussian process regression models that are used for predicting capacity at early/mid-life cycles for the test cells (see blue dots with error bars in Fig 1b). Finally, these predicted data-points are used in addition to the actual early cycle data capacity fade to construct the Bayesian regression trajectory for the test cell s. Notably. We note that these two methods are complementary and can be combined with each other. We evaluate the performance of our proposed method on an testing open-source dataset from Iowa State University and Iowa Lakes Community College (ISU-ILCC). This dataset comprises of 251 nickel-manganese-cobalt/graphite Lithium-ion cells that are cycled under 63 different conditions. We compute the mean average percentage error (MAPE) and negative log predictive density (NLPD) to quantify the efficacy of our method. Our initial findings suggest that, when only few observations are available, for test cells, when using only Bayesian regression with uninformed priors, a power law functional provides the most accurate predictions. with very few data points. However, asHowever, a the number of data points increases, a twin sigmoidal function becomes more accurate as the number of observations further increases. We also find that using as little as 10% of the data set towards generating data-driven priors can lead to significant improvement in prediction accuracy when using early cycle data. Lastly, we found that augmenting early-cycle data with Gaussian process-predicted capacity data for Bayesian regression greatly improves the prediction accuracy. We will present a comprehensive comparison of our methods to other methods available in the literature and apply this method to additional battery datasets.

42 ENGINEERING↗

Data Format and Descriptions for the Alabama Carbon Storage: Data Sharing and Engagement Project

The Alabama Carbon Storage: Data Sharing and Engagement (ACS-DSE) project seeks to develop publicly accessible geologic carbon storage models and data across the southern Gulf Coastal Plain of Alabama. The public online platform developed for this project will include geologic, geophysical, infrastructure, and other relevant datasets and geologic models of the study area. Datasets, model surfaces (e.g. structural contour maps, isolith maps, porosity maps), and infrastructure data (e.g. offshore pipelines, field boundaries) will be downloadable in commonly used file formats. The anticipated primary geologic datasets are well headers, formation tops, average reservoir properties, and core analyses; these will be available as commaseparated values (CSV) text files and MS Excel workbooks. Geophysical logs will be available in Log ASCII Standard (LAS) file format. Modeled surfaces, such as structure contour maps, will be available in ArcGIS formats and text files. Infrastructure data will be available as ArcGIS shapefiles. This document provides information on the data sources and attributes of the datasets.

01 COAL, LIGNITE, AND PEAT↗

MCPC Friction Stir Welding (FSW) Process Data

Processing parameters and machine log data for the MCPC LDRD Agile investment is collected material samples processed. This dataset captures the selected processing parameters, machine logs captured during material processing, and descriptions of how characterization samples were extracted from processed plates of material. The collect characterization data is captured in other datasets.

316 Stainless Steel↗

Cybersecurity Certification Requirements for Distributed Energy Resources: A Survey of SunSpec Alliance Standards

This survey paper explores the cybersecurity certification requirements defined by the SunSpec Alliance for Distributed Energy Resource (DER) devices, focusing on aspects such as software updates, device communications, authentication mechanisms, device security, logging, and test procedures. The SunSpec cybersecurity standards mandate support for remote and automated software updates, secure communication protocols, stringent authentication practices, and robust logging mechanisms to ensure operational integrity. Furthermore, the paper discusses the implementation of the SAE J3072 standard using the IEEE 2030.5 protocol, emphasizing the secure interactions between electric vehicle supply equipment (EVSE) and plug-in electric vehicles (PEVs) for functionalities like vehicle-to-grid (V2G) capabilities. This research also examines the SunSpec Modbus standard, which enhances the interoperability among DER system components, facilitating compliance with grid interconnection standards. This paper also analyzes the existing SunSpec Device Information Models, which standardize data exchange formats for DER systems across communication interfaces. Finally, this paper concludes with a detailed discussion of the energy storage cybersecurity specification and the blockchain cybersecurity requirements as proposed by SunSpec Alliance.

Tsikteris, Sean (ORCID:0009000524202250)↗

Red Noise–based False Alarm Thresholds for Astrophysical Periodograms via Whittle’s Approximation to the Likelihood

Astronomers who search for periodic signals using Lomb–Scargle periodograms rely on false alarm level (FAL) estimates to identify statistically significant peaks. Although FALs are often calculated from white noise models, many astronomical time series suffer from red noise. Prewhitening is a statistical technique in which a continuum model is subtracted from the log power spectrum estimate, after which the observer can proceed with a white-noise treatment. Here we present a prewhitening-based method of calculating frequency-dependent FALs. We fit power laws and autoregressive models of order 1 to each Lomb–Scargle periodogram by minimizing the Whittle approximation to the negative log-likelihood (NLL), then calculate FALs based on the best-fit model power spectrum. Our technique is a novel extension of the Whittle NLL to datasets with uneven time sampling. We demonstrate FAL calculations using observations of α Cen B, GJ 581, HD 192310, synthetic data from the radial velocity (RV) fitting challenge, and Kepler observations of a differential rotator. The Kepler data analysis shows that only true rotation signals are detected by red noise FALs, while white noise FALs suggest all spurious peaks in the low-frequency range are significant. A high-frequency sinusoid injected into α Cen B logR$'$ HK observations exceeds the 1% red noise FAL despite having only 8.9% of the power of the dominant rotation signal. In a periodogram of HD 192310 RVs, peaks associated with differential rotation and planets are detected against the 5% red noise FAL without iterative model fitting or subtraction. The software for calculating red noise–based FALs is available on GitHub.

Astrostatistics (1882)↗

DESI Massive Poststarburst Galaxies at z ~ 1.2 Have Compact Structures and Dense Cores

Poststarburst galaxies (PSBs) are young quiescent galaxies that have recently experienced a rapid decrease in star formation, allowing us to probe the fast-quenching period of galaxy evolution. In this work, we obtained Hubble Space Telescope (HST)/WFC3 F110W imaging to measure the sizes of 171 massive (log(M $\ast$ /M ⊙ ) ~ 11) spectroscopically identified PSBs at 1 < z 1.3 selected from the DESI Survey Validation luminous red galaxy sample. This statistical sample constitutes an order of magnitude increase from the ~20 PSBs with space-based imaging and deep spectroscopy. We perform structural fitting of the target galaxies with pysersic and compare them to quiescent and star-forming galaxies in the 3D-HST survey. We find that these PSBs are more compact than the general population of quiescent galaxies, lying systematically ~0.1 dex below the established size–mass relation. However, their central surface mass densities are similar to those of their quiescent counterparts (log(Σ 1kpc /(M ⊙ kpc -2 ))~10.1). These findings are easily reconciled by later ex situ growth via minor mergers or a slight progenitor bias. These PSBs are round in projection (b/a median ~ 0.8), suggesting that they are primarily spheroids, not disks, in 3D. We find no correlation between the time since quenching and light-weighted PSB sizes or central densities. This disfavors apparent structural growth due to the fading of centralized starbursts in this galaxy population. Instead, we posit that the fast quenching of massive galaxies at this epoch occurs preferentially in galaxies with preexisting compact structures.

79 ASTRONOMY AND ASTROPHYSICS↗

The SAGA Survey. VI. The Size–Mass Relation for Low-mass Galaxies Across Environments

We investigate how Milky Way (MW)–like environments influence the sizes and structural properties of low mass galaxies by comparing satellites of MW analogs from the Satellites Around Galactic Analogs (SAGA) Survey with two control samples: an environmentally agnostic population from the SAGA background sample and isolated galaxies from the Sloan Digital Sky Survey NASA-Sloan Atlas. All sizes and structural parameters are measured uniformly using pysersic to ensure consistency across samples. We find the half-light sizes of SAGA satellites are systematically larger than those of isolated galaxies, with the magnitude of the offset ranging from 0.05 to 0.12 dex (10%–24%) depending on the comparison sample and completeness cuts. This corresponds to physical size differences between 85 and 200 pc at log 10 (M$_{{\star}}$/M ⊙ ) = 7.5 and 220–960 pc at log 10 (M$_{{\star}}$/M ⊙ ) = 10. This offset persists among star-forming galaxies, suggesting that environment can influence the structure of low-mass galaxies even before it impacts quenching. The intrinsic scatter in the size–mass relation is lower for SAGA satellites than isolated galaxies, and the Sérsic index distributions of satellites and isolated galaxies are similar. In comparison to star-forming satellites, quenched SAGA satellites have a slightly shallower size–mass relation and rounder morphologies at low mass, suggesting that quenching is accompanied by structural transformation and that the processes responsible differ between low- and high-mass satellites. Our results show that environmental processes can imprint measurable structural differences on satellites in MW-mass halos.

Asali, Yasmeen [Yale Univ., New Haven, CT (United ↗

Determining Stellar Elemental Abundances from DESI Spectra with the Data-driven Payne

Abstract Stellar abundances for a large number of stars provide key information for the study of Galactic formation history. Large spectroscopic surveys such as the Dark Energy Spectroscopic Instrument (DESI) and LAMOST take median-to-low-resolution (R≲ 5000) spectra in the full optical wavelength range for millions of stars. However, the line-blending effect in these spectra causes great challenges for elemental abundance determination. Here we employDD-Payne, a data-driven method regularized by differential spectra from stellar physical models, to the DESI early data release spectra for stellar abundance determination. Our implementation delivers 15 labels, including effective temperatureT eff , surface gravity log g , microturbulence velocityv mic , and the abundances for 12 individual elements, namely C, N, O, Mg, Al, Si, Ca, Ti, Cr, Mn, Fe, and Ni. Given a spectral signal-to-noise ratio of 100 per pixel, the internal precisions of the label estimates are about 20 K forT eff , 0.05 dex for log g , and 0.05 dex for most elemental abundances. These results agree with the theoretical limits from the Crámer–Rao bound calculation within a factor of 2. The majority of the accreted halo stars contributed by the Gaia–Enceladus–Sausage are discernible from the disk and in situ halo populations in the resultant [Mg/Fe]–[Fe/H] and [Al/Fe]–[Fe/H] abundance spaces. We also provide distance and orbital parameters for the sample stars, which spread over a distance out to ∼100 kpc. The DESI sample has a significantly higher fraction of distant (or metal-poor) stars than the other existing spectroscopic surveys, making it a powerful data set for studying the Galactic outskirts. The catalog is publicly available.

Astronomy & Astrophysics↗

Score-based deterministic density sampling

We propose a deterministic sampling framework using Score-Based Transport Modeling for sampling an unnormalized target density π given only its score ∇ log π. Our method approximates the Wasserstein gradient flow on KL($f_t$∥π) by learning the time-varying score ∇ log $f_t$ on the fly using score matching. While having the same marginal distribution as Langevin dynamics, our method produces smooth deterministic trajectories, resulting in monotone noise-free convergence. We prove that our method dissipates relative entropy at the same rate as the exact gradient flow, provided sufficient training. Numerical experiments validate our theoretical findings: our method converges at the optimal rate, has smooth trajectories, and is often more sample efficient than its stochastic counterpart. Experiments on high-dimensional image data show that our method produces high-quality generations in as few as 15 steps and exhibits natural exploratory behavior. The memory and runtime scale linearly in the sample size.

97 MATHEMATICS AND COMPUTING↗

Concentration-Discharge Relationships in the Six Largest Arctic Rivers, 2003-2019

This dataset provides the results of the analysis of the relationship of dissolved analyte concentrations and river discharges in the six largest Arctic rivers across the global panarctic region (see Figure 1 in documentation file *.pdf). Long-term measurements of dissolved analyte concentrations and river discharge have been collected for each of the Kolyma, Lena, Mackenzie, Ob, Yenisey, and Yukon rivers by the Arctic Great Rivers Observatory (ArcticGRO) project from ~2003-present (Shiklomanov, 2021). The relationship of dissolved analyte concentrations and discharges in each river was characterized by statistical analysis of the slope of the log(concentration) vs log(discharge) (b), the coefficient of variation ratio (CVc/CVq), the 2.5% and 97.5% confidence intervals of b, and assigning a chemostatic, flushing, diluting, or non-systematic behavior category according to Koger (2018). The summary of these analyses for all six rivers is provided in one .csv file. The concentrations of 20 dissolved analytes and discharge measurement data for the individual Kolyma, Lena, Mackenzie, Ob, Yenisey, and Yukon rivers are also provided with this dataset. There are seven *.csv files; one for each river plus the statistical summary. These public ArcticGRO data at "https://www.arcticgreatrivers.org" (Shiklomanov, 2021) were downloaded on Feb 13, 2020, but each river has different measurement dates over the sampling and analysis period. The ArcticGRO metadata document (*.pdf) downloaded on Feb 13, 2020 is also included in this dataset. The Next-Generation Ecosystem Experiments: Arctic (NGEE Arctic), was a research effort to reduce uncertainty in Earth System Models by developing a predictive understanding of carbon-rich Arctic ecosystems and feedbacks to climate. NGEE Arctic was supported by the Department of Energy's Office of Biological and Environmental Research. The NGEE Arctic project had two field research sites: 1) located within the Arctic polygonal tundra coastal region on the Barrow Environmental Observatory (BEO) and the North Slope near Utqiagvik (Barrow), Alaska and 2) multiple areas on the discontinuous permafrost region of the Seward Peninsula north of Nome, Alaska. Through observations, experiments, and synthesis with existing datasets, NGEE Arctic provided an enhanced knowledge base for multi-scale modeling and contributed to improved process representation at global pan-Arctic scales within the Department of Energy's Earth system Model (the Energy Exascale Earth System Model, or E3SM), and specifically within the E3SM Land Model component (ELM).

54 ENVIRONMENTAL SCIENCES↗

Streaming Matching and Edge Cover in Practice

Graph algorithms with polynomial space and time requirements often become infeasible for massive graphs with billions of edges or more. State-of-the-art approaches therefore employ approximate serial, parallel, and distributed algorithms to tackle these challenges. However, such approaches require storing the entire graph in memory and thus need access to costly computing resources such as clusters and supercomputers. In this paper, we present practical streaming approaches for solving massive graph problems using limited memory for two prototypical graph problems: maximum weighted matching and minimum weighted edge cover. For matching, we conduct a thorough computational study on two of the semi-streaming algorithms including a recent breakthrough result that achieves a $1/(2+\varepsilon)$-approximation of the weight while using $O( n \log W /\epsilon)$ memory (here $n$ is the number of vertices and $W$ is the maximum edge weight), designed by Paz and Schwartzman [SODA, 2017]. Empirically, we show that the semi-streaming algorithms produce matchings whose weight is close to the best $1/2$-approximate offline algorithm while requiring less time and an order-of-magnitude less memory. For minimum weighted edge cover, we develop three novel semi-streaming algorithms. Two of these algorithms require a single pass through the input graph, require $O(n \log n)$ memory, and provide a 2-approximation guarantee on the objective. We also leverage a relationship between approximate maximum weighted matching and approximate minimum weighted edge cover to develop a two-pass $3/2+\epsilon$-approximate algorithm with the memory requirement of Paz and Schwartzman's semi-streaming matching algorithm. These streaming approaches are compared against the state-of-the-art 3/2-approximate offline algorithm. The semi-streaming matching and the novel edge cover algorithms proposed in this paper can process graphs with several billions of edges in under 30 minutes using 6 GB of memory, which is at least an order of magnitude improvement from the offline (non-streaming) algorithms. For the largest graph, the best alternative offline parallel approximation algorithm (GPA+ROMA) could not finish in three hours even while employing hundreds of processors and 1 TB of memory. We also demonstrate an application of the semi-streaming algorithm by computing a matching using linearly bounded memory on item intersection graphs derived from three machine learning datasets, whereas the existing offline algorithms could not complete on one of these datasets since their memory requirements exceeded 1TB.

Ferdous, S M.↗

Secure hierarchical processing using a secure ledger

Disclosed is a system and method for processing data using blockchain technology. The system includes a memory having programmable instructions stored thereon that, when executed by a processor, cause the system to: authenticate one or more sensors in anticipation of receiving component data; receive component data, upon successful authentication; store the component data locally or to a cloud-based server and/or calculate a root value for the component data; store or embed the root value with the stored component data; condense the component data and link the condensed component data to the stored component data via the root value. The system further includes instructions to log the condensed data, including the root value, to a ledger, and to identify a tag or transaction id corresponding to the logging event for subsequent retrieval of the condensed data using the tag or transaction id.

Zhao, Wenbing↗

A DECADE of dwarfs: first detection of weak lensing around spectroscopically confirmed low-mass galaxies

We present the first detection of weak gravitational lensing around spectroscopically confirmed dwarf galaxies, using the large overlap between DESI DR1 spectroscopic data and DECADE/DES weak lensing catalogs. A clean dwarf galaxy sample with well-defined redshift and stellar mass cuts enables excess surface mass density measurements in two stellar mass bins ($\log \rm{M}_*=[8.2, 9.2]~M_\odot$ and $\log \rm{M}_*=[9.2, 10.2]~M_\odot$), with signal-to-noise ratios of $5.6$ and $12.4$ respectively. This signal-to-noise drops to $4.5$ and $9.2$ respectively for measurements without applying individual inverse probability (IIP) weights, which mitigates fiber incompleteness from DESI's targeting. The measurements are robust against variations in stellar mass estimates, photometric shredding, and lensing calibration systematics. Using a simulation-based modeling framework with stellar mass function priors, we constrain the stellar mass-halo mass relation and find a satellite fraction of $\simeq 0.3$, which is higher than previous photometric studies but $1.5σ$ lower than $Λ$CDM predictions. We find that IIP weights have a significant impact on lensing measurements and can change the inferred $f_{\rm{sat}}$ by a factor of two, highlighting the need for accurate fiber incompleteness corrections for dwarf galaxy samples. Our results open a new observational window into the galaxy-halo connection at low masses, showing that future massively multiplexed spectroscopic observations and weak lensing data will enable stringent tests of galaxy formation models and $Λ$CDM predictions.

To, Chun-Hao [Chicago U., Astron. Astrophys. Ctr.;↗