Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Hierarchical Data Format”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Hyperspectral remote sensing-based plant community map for region around NGEE-Arctic intensive research watersheds at Seward Peninsula, Alaska, 2017-2019

Using airborne hyperspectral remote sensing data from NASA Airborne Visible-Infrared Imaging Spectrometer- Next Generation (AVIRIS-NG) platforms in a region near NGEE-Arctic intensive watersheds at Seward peninsula of Alaska, high resolution (5m) maps of plant community distribution were developed and included in this data collected. AVIRIS-NG data collected over 2017-2019 period were used to develop deep neural networks, trained using vegetation plot observations collected at NGEE-Arctic watersheds at Kougarok, Council and Teller. A hierarchical vegetation classification scheme consisting of six classes at Level I, and 16 classes at Level II contained in two .txt files were used to developed the plant community maps for the region. Two geospatial raster data files (.tif) at both thematic levels are shared in this data collection. Data files in this collection use Alaska Albers Equal Area projection. Readme files available in three formats (*.html, *.md, *.pdf) and one *.png visualization map.The Next-Generation Ecosystem Experiments: Arctic (NGEE Arctic), was a research effort to reduce uncertainty in Earth System Models by developing a predictive understanding of carbon-rich Arctic ecosystems and feedbacks to climate. NGEE Arctic was supported by the Department of Energy's Office of Biological and Environmental Research.The NGEE Arctic project had two field research sites: 1) located within the Arctic polygonal tundra coastal region on the Barrow Environmental Observatory (BEO) and the North Slope near Utqiagvik (Barrow), Alaska and 2) multiple areas on the discontinuous permafrost region of the Seward Peninsula north of Nome, Alaska.Through observations, experiments, and synthesis with existing datasets, NGEE Arctic provided an enhanced knowledge base for multi-scale modeling and contributed to improved process representation at global pan-Arctic scales within the Department of Energy's Earth system Model (the Energy Exascale Earth System Model, or E3SM), and specifically within the E3SM Land Model component (ELM).

54 ENVIRONMENTAL SCIENCES↗

Non-linear hydrologic organization

We revisit three variants of the well-known Stommel diagrams that have been used to summarize knowledge of characteristic scales in time and space of some important hydrologic phenomena and modified these diagrams focusing on spatiotemporal scaling analyses of the underlying hydrologic processes. In the present paper we focus on soil formation, vegetation growth, and drainage network organization. We use existing scaling relationships for vegetation growth and soil formation, both of which refer to the same fundamental length and timescales defining flow rates at the pore scale but different powers of the power law relating time and space. The principle of a hierarchical organization of optimal subsurface flow paths could underlie both root lateral spread (RLS) of vegetation and drainage basin organization. To assess the applicability of scaling, and to extend the Stommel diagrams, data for soil depth, vegetation root lateral spread, and drainage basin length have been accessed. The new data considered here include timescales out to 150 Myr that correspond to depths of up to 240 m and horizontal length scales up to 6400 km and probe the limits of drainage basin development in time, depth, and horizontal extent.

58 GEOSCIENCES↗

A Kinematic Perspective on the Formation Process of the Stellar Groups in the Rosette Nebula

Stellar kinematics is a powerful tool for understanding the formation process of stellar associations. Here, we present a kinematic study of the young stellar population in the Rosette nebula using recent Gaia data and high-resolution spectra. We first isolate member candidates using the published mid-infrared photometric data and the list of X-ray sources. A total of 403 stars with similar parallaxes and proper motions are finally selected as members. The spatial distribution of the members shows that this star-forming region is highly substructured. The young open cluster NGC 2244 in the center of the nebula has a pattern of radial expansion and rotation. We discuss its implication on the cluster formation, e.g., monolithic cold collapse or hierarchical assembly. On the other hand, we also investigate three groups located around the border of the H ii bubble. The western group seems to be spatially correlated with the adjacent gas structure, but their kinematics is not associated with that of the gas. The southern group does not show any systematic motion relative to NGC 2244. These two groups might be spontaneously formed in filaments of a turbulent cloud. The eastern group is spatially and kinematically associated with the gas pillar receding away from NGC 2244. This group might be formed by feedback from massive stars in NGC 2244. Our results suggest that the stellar population in the Rosette Nebula may form through three different processes: the expansion of stellar clusters, hierarchical star formation in turbulent clouds, and feedback-driven star formation.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Biolink Model: A universal schema for knowledge graphs in clinical, biomedical, and translational science

Abstract Within clinical, biomedical, and translational science, an increasing number of projects are adopting graphs for knowledge representation. Graph‐based data models elucidate the interconnectedness among core biomedical concepts, enable data structures to be easily updated, and support intuitive queries, visualizations, and inference algorithms. However, knowledge discovery across these “knowledge graphs” (KGs) has remained difficult. Data set heterogeneity and complexity; the proliferation of ad hoc data formats; poor compliance with guidelines on findability, accessibility, interoperability, and reusability; and, in particular, the lack of a universally accepted, open‐access model for standardization across biomedical KGs has left the task of reconciling data sources to downstream consumers. Biolink Model is an open‐source data model that can be used to formalize the relationships between data structures in translational science. It incorporates object‐oriented classification and graph‐oriented features. The core of the model is a set of hierarchical, interconnected classes (or categories) and relationships between them (or predicates) representing biomedical entities such as gene, disease, chemical, anatomic structure, and phenotype. The model provides class and edge attributes and associations that guide how entities should relate to one another. Here, we highlight the need for a standardized data model for KGs, describe Biolink Model, and compare it with other models. We demonstrate the utility of Biolink Model in various initiatives, including the Biomedical Data Translator Consortium and the Monarch Initiative, and show how it has supported easier integration and interoperability of biomedical KGs, bringing together knowledge from multiple sources and helping to realize the goals of translational science.

60 APPLIED LIFE SCIENCES↗

Bit-GraphBLAS: Bit-Level Optimizations of Matrix-Centric Graph Processing on GPU

In the graph data structure like adjacency matrix, the connectivity of two nodes can be sufficiently represented using only 1 bit, but they are generally treated as 32-bit full-precision in state-of-the-art graph frameworks to adopt common sparse format such as CSR. Meanwhile, bit-level parallelism has recently be explored to have high-performance potential and low storage requirement on GPUs with dense bit-tiles. To fill the gap, our solution is a hierarchical storage format that contains the bit-indexing base and dense bit-tile units. Inherently, the granularity of the bit-tile is an essential factor in achieving both storage compression and GPU parallelism. How to find a sweet spot that trades off between avoiding sparsity and exploiting is comprehensively researched in this work. In the experiment, we evaluate the proposed storage format and algorithms on modern generation GPUs, including Pascal and Volta, to figure out critical software co-designs in conjunction with existing hardware-specific optimization.

Chen, Jou-An↗

Robust Carbon Dioxide Plume Imaging Using Joint Tomographic Inversion of Seismic Onset Time and Distributed Pressure and Temperature Measurements (Final Report)

We develop and demonstrate rapid and cost-effective methodologies for spatiotemporal tracking of CO2 plumes during geologic sequestration using joint inversion of seismic data and distributed pressure and temperature measurements. Key elements of our methodology are: (a) a computationally efficient approach to pressure and temperature propagation, (b) analysis of time lapse seismic data using a novel ‘seismic onset time’ approach to detect fluid front propagation, and (c) data assimilation and uncertainty assessment via joint inversion of pressure, temperature and time lapse seismic data, and (d) validating the numerical tomographic inversion using a CO2 injection demonstration projects, specifically data collected from the from the Petra Nova Parish Holdings CCUS project in the West Ranch Field, Texas and the Chester-16 reef CO2 injection site in Northern Michigan which is part of the DOE Midwestern Carbon Sequestration Project. The research team is led by Texas A&M University and includes Battelle as a subcontractor with support from Shell, Anadarko, Chevron and JX Nippon. A carbon dioxide (CO2) water-alternating-gas (WAG) pilot was conducted to gain insights into tertiary oil recovery potential via CO2 flood in the West Ranch Field as part of the Petra Nova project, the world’s largest post-combustion CO2 capture and utilization initiative. With a fluvial formation geology and large contrasts in permeability, this is a challenging and novel application of CO2 enhanced oil recovery (EOR). We build a predictive dynamic model of the subsurface that incorporates the multiphase and compositional data acquired during the pilot operation. The calibrated model is used for the carbon dioxide plume imaging. The study began with an initialization of the pilot sector model extracted from a calibrated full-field model. The pilot model calibration follows a two-step hierarchical workflow. First, we performed a large-scale update of the permeability distribution by integrating available bottomhole pressure and multiphase production data. In the second step, local permeability field is fine-tuned using a streamline-based method to match CO2 breakthrough times at the producers. The predictive capability of the calibrated model was verified through two blind validation tests: (1) the model showed good agreement with saturation logs acquired at two observation wells; and (2) the model reproduced the CO2 recovery as a fraction of the injected CO2. The use of seismic onset times has shown great promise for integrating near-continuous seismic surveys for updating geologic models. In this study, we analyze the impact of seismic survey frequency on the onset time approach aiming to extend the application of onset time to infrequent seismic surveys. In addition, we quantitatively examine the nonlinearity of the onset time method and compare it to the commonly used amplitude inversion method. We carry out a sensitivity analysis of seismic survey frequency based on the complete seismic survey data (over 175 surveys) of steam injection in a heavy oil reservoir (Peace River Unit) in Canada. Our results show that an adequate onset time map can be obtained from the infrequent seismic surveys by interpolation between seismic surveys as long as there is no change in the dominant underlying physics between the successive surveys. The study also shows that nonlinearity of the onset time method can be -smaller than that of the amplitude inversion method by several orders of magnitude. Application to the Brugge benchmark case shows that the onset time method obtains comparable permeability update as the traditional seismic amplitude inversion method with faster computation and improved convergence characteristics. We extend the streamline-based data integration approach to incorporate distributed temperature sensor (DTS) data using the concept of thermal tracer travel time. Then, a hierarchical workflow composed of evolutionary and streamline methods is employed to jointly history match the DTS and pressure data. Finally, CO2 saturation and streamline maps are used to visualize the CO2 plume movement during the sequestration process. The hierarchical workflow is applied to a carbon sequestration project in a carbonate reef reservoir within the Northern Niagaran Pinnacle Reef Trend in Michigan, USA. The monitoring data set consists of distributed temperature sensing (DTS) data acquired at the injection well and a monitoring well, flowing bottom-hole pressure data at the injection well, and time-lapse pressure measurements at several locations along the monitoring well. The history matching results indicate that the CO2 movement is mostly restricted to the intended zones of injection which is consistent with an independent warm-back analysis of the temperature data. In addition to employing simulation models and inverse methods for CO2 plume imaging, we also initialized a data-driven technology for detecting inter-well connectivity based on production and pressure data. Our machine-learning framework is built on the statistical recurrent unit (SRU) model and interprets well-based injection/production data into inter-well connectivity without relying on a geologic model. We test it on synthetic and field-scale CO2 EOR projects utilizing the water-alternating-gas (WAG) process. The validation of the proposed data-driven inter-well connectivity assessment is performed using synthetic data from simulation models where inter-well connectivity can be easily measured using the streamline-based flux allocation. The SRU model is shown to offer excellent prediction performance on the synthetic case. Despite significant measurement noise and frequent well shut-ins imposed in the field-scale case, the SRU model offers good prediction accuracy, the overall relative error of the phase production rates at most producers ranges from 10% to 30%. It is shown that the dominant connections identified by the data-driven method and streamline method are in close agreement. Texas A&M University, the lead organization in the project, was primarily responsible for the development of tomographic approaches for CO2 plume mapping in conjunction with distributed pressure, temperature and seismic onset time data. Battelle, as a subcontractor, was primarily responsible for the development of analytical and empirical methods for analyzing transient injection rate and pressure data from point/line sources such as injection and monitoring wells. An additional area of emphasis for Battelle was the use of machine learning for such tasks as inferring reservoir connectivity information from injection-production data, and identifying variable importance for machine learning-based proxy models developed from full-physics simulations. The two organizations also collaborated on the application of the tomographic inversion methodology for a field data set.

02 PETROLEUM↗

AMReX v2024

The software framework, AMReX, supports the development of block-structured adaptive mesh refinement (AMR) algorithms for solving systems of partial differential equations. AMR reduces the computational cost and memory footprint compared to a uniform mesh while preserving the essential local descriptions of different physical processes in complex multiphysics algorithms. AMR uses a hierarchical representation of the solution at multiple levels of resolution where the solution on each level is defined on the union of data containers at that resolution. These data containers, which represent the solution over a logically rectangular subregion of the domain, can contain field data defined on a mesh, Lagrangian particles or combinations of both. In addition to these basic data types, AMReX supports a multilevel embedded boundary representation of complex geometry; linear solvers for cell-centered and nodal data; asynchronous I/O in a native format readable by ParaView, VisIt and yt; and interfaces to hypre and PETSc solvers. AMReX enables applications to run on distributed memory architectures with multicore CPUs and with GPU accelerators. AMReX uses a lightweight abstraction layer that effectively hides the details of the architecture from the application. The framework currently supports CUDA, HIP and SYCL for GPU acceleration and OpenMP for multi-core CPU architectures.

Almgren, Ann↗

ARMADA. I. Triple Companions Detected in B-type Binaries α Del and ν Gem

Ground-based optical long-baseline interferometry has the power to measure the orbits of close binary systems at ∼10 μas precision. This precision makes it possible to detect “wobbles” in the binary motion due to the gravitational pull from additional short-period companions. We started the ARrangement for Micro-Arcsecond Differential Astrometry (ARMADA) survey with the Michigan Infra-Red Combiner (MIRC)/MIRC-X instrument at the Center for High Angular Resoloution Astronomy (CHARA) array for the purpose of detecting giant planets and stellar companions orbiting individual stars in binary systems. We describe our observations for the survey, and introduce the wavelength calibration scheme that delivers precision at the tens of microarcseconds level for <02 binaries. We test our instrument performance on a known triple system, κ Peg, and show that our survey is delivering a factor of 10 better precision than previous similar surveys. We present astrometric detections of tertiary components to two B-type binaries: a 30 day companion to α Del, and a 50 day companion to ν Gem. We also collected radial velocity data for α Del with the Tennessee State University Automated Spectroscopic Telescope at Fairborn Observatory. We are able to measure the orbits and masses of all three components in these systems. We find that the previously published radial velocity orbit for the inner pair of ν Gem is not consistent with our visual orbit. The precision achieved for these orbits suggests that our ARMADA survey will be successful at discovering new compact triple systems to A/B-type binary systems, leading to better statistics of hierarchical system architectures and formation history.

47 OTHER INSTRUMENTATION↗

Discovery of a Candidate Binary Supermassive Black Hole in a Periodic Quasar from Circumbinary Accretion Variability

Binary supermassive black holes (BSBHs) are expected to be a generic byproduct from hierarchical galaxy formation. The final coalescence of BSBHs is thought to be the loudest gravitational wave (GW) siren, yet no confirmed BSBH is known in the GW-dominated regime. While periodic quasars have been proposed as BSBH candidates, the physical origin of the periodicity has been largely uncertain. Here we report discovery of a periodicity (P=1607±7 days) at 99.95% significance (with a global p-value of ~10 –3 accounting for the look elsewhere effect) in the optical light curves of a redshift 1.53 quasar, SDSS J025214.67–002813.7. Combining archival Sloan Digital Sky Survey data with new, sensitive imaging from the Dark Energy Survey, the total ~20-yr time baseline spans ~4.6 cycles of the observed 4.4-yr (restframe 1.7-yr) periodicity. The light curves are best fit by a bursty model predicted by hydrodynamic simulations of circumbinary accretion disks. The periodicity is likely caused by accretion rate modulation by a milli-parsec BSBH emitting GWs, dynamically coupled to the circumbinary accretion disk. A bursty hydrodynamic variability model is statistically preferred over a smooth, sinusoidal model expected from relativistic Doppler boost, a kinematic effect proposed for PG1302–102. Furthermore, the frequency dependence of the variability amplitudes disfavors Doppler boost, lending independent support to the circumbinary accretion variability hypothesis. Given our detection rate of one BSBH candidate from circumbinary accretion variability out of 625 quasars, it suggests that future large, sensitive synoptic surveys such as the Vera C. Rubin Observatory Legacy Survey of Space and Time may be able to detect hundreds to thousands of candidate BSBHs from circumbinary accretion with direct implications for Laser Interferometer Space Antenna.

79 ASTRONOMY AND ASTROPHYSICS↗

Understanding the origin of early-type dwarfs: the spectrophotometric study of CGCG014−074

ABSTRACT Early-type dwarf galaxies constitute a prevalent population in the central regions of rich groups and clusters in the local Universe. These low-luminosity and low-mass stellar systems play a fundamental role in the assembly of the luminous galaxies observed today, according to the Lambda cold dark matter hierarchical theory. The origin of early-type dwarfs has been linked to the transformation of disc galaxies interacting with the intracluster medium, especially in dense environments. However, the existence of low-luminosity early-type galaxies in low-density environments presents a challenge to this scenario. This study presents a comprehensive photometric and spectroscopic analysis of the early-type dwarf galaxy CGCG014−074 using deep Gemini GMOS (Gemini Multi-Object Spectrograph) data, focusing on its peculiarities and evolutionary implications. CGCG014−074 exhibits distinct features, including a rotating inner disc, an extended stellar formation with a quiescent phase since about 2 Gyr ago, and the presence of boxy isophotes. From the kinematic analysis, we confirm CGCG014−074 as a nucleated early-type dwarf galaxy with embedded disc. The study of its stellar population parameters using different methods provides significant insights into the galaxy’s evolutionary history. These results show an old and metal-poor nucleus (${\sim}9.3$ Gyr and $\mathrm{[Z/H]}\sim -0.84$ dex), while the stellar disc is younger (${\sim}4.4$ Gyr) with a higher metallicity ($\mathrm{[Z/H]}\sim -0.40$ dex). These distinctive features collectively position CGCG014−074 as a likely building block galaxy that has evolved passively throughout its history.

Astronomy & Astrophysics↗

HarDWR - Harmonized Water Rights Records

A dataset within the Harmonized Database of Western U.S. Water Rights (HarDWR). For a detailed description of the database, please see the meta-record v2.0. Changelog v2.0 - Recalculated based on data sourced from WestDAAT - Changed using a Site ID column to identify unique records to using aa combination of Site ID and Allocation ID - Removed the Water Management Area (WMA) column from the harmonized records. The replacement is a separate file which stores the relationship between allocations and WMAs. This allows for allocations to contribute to water right amounts to multiple WMAs during the subsequent cumulative process. - Added a column describing a water rights legal status - Added "Unspecified" was a water source category - Added an acre-foot (AF) column - Added a column for the classification of the right's owner v1.02 - Added a .RData file to the dataset as a convenience for anyone exploring our code. This is an internal file, and the one referenced in analysis scripts as the data objects are already in R data objects. v1.01 - Updated the names of each file with an ID number less than 3 digits to include leading 0s v1.0 - Initial public release Description Here we present an updated database of Western U.S. water right records. This database provides consistent unique identifiers for each water right record, and a consistent categorization scheme that puts each water right record into one of seven broad use categories. These data were instrumental in conducting a study of the multi-sector dynamics of inter-sectoral water allocation changes though water markets (Grogan et al., *in review*). Specifically, the data were formatted for use as input to a process-based hydrologic model, Water Balance Model (WBM), with a water rights module (Grogan et al., *in review*). While this specific study motivated the development of the database presented here, water management in the U.S. West is a rich area of study (e.g., Anderson and Woosly, 2005; Tidwell, 2014; Null and Prudencio, 2016; Carney et al., 2021) so releasing this database publicly with documentation and usage notes will enable other researchers to do further work on water management in the U.S. West. We produced the water rights database presented here in four main steps: (1) data collection, (2) data quality control, (3) data harmonization, and (4) generation of cumulative water rights curves. Each of steps (1)-(3) had to be completed in order to produce (4), the final product that was used in the modeling exercise in Grogan et al. (*in review*). All data in each step is associated with a spatial unit called a Water Management Area (WMA), which is the unit of water right administration utilized by the state in which the right came from. Steps (2) and (3) required use to make assumptions and interpretation, and to remove records from the raw data collection. We describe each of these assumptions and interpretations below so that other researchers can choose to implement alternative assumptions an interpretation as fits their research aims. Motivation for Changing Data Sources The most significant change has been a switch from collecting the raw water rights directly from each state to using the water rights records presented in WestDAAT, a product of the Water Data Exchange (WaDE) Program under the Western States Water Council (WSWC). One of the main reasons for this is that each state of interest is a member of the WSWC, meaning that WaDE is partially funded by these states, as well as many universities. As WestDAAT is also a database with consistent categorization, it has allowed us to spend less time on data collection and quality control and more time on answering research questions. This has included records from water right sources we had previously not known about when creating v1.0 of this database. The only major downside to utilizing the WestDAAT records as our raw data is that further updates are tied to when WestDAAT is updated, as some states update their public water right records daily. However, as our focus is on cumulative water amounts at the regional scale, it is unlikely most records updates would have a significant effect on our results. The structure of WestDAAT led to several important changes to how HarWR is formatted. The most significant change is that WaDE has calculated a field known as `SiteUUID`, which is a unique identifier for the Point of Diversion (POD), or where the water is drawn from. This separate from `AllocationNativeID`, which is the identifier for the allocation of water, or the amount of water associated with the water right. It should be noted that it is possible for a single site to have multiple allocations associated with it and for an allocation to be able to be extracted from multiple sites. The site-allocation structure has allowed us to adapt a more consistent, and hopefully more realistic, approach in organizing the water right records than we had with HarDWR v1.0. This was incredibly helpful as the raw data from many states had multiple water uses within a single field within a single row of their raw data, and it was not always clear if the first water use was the most important, or simply first alphabetically. WestDAAT has already addressed this data quality issue. Furthermore, with v1.0, when there were multiple records with the same water right ID, we selected the largest volume or flow amount and disregarded the rest. As WestDAAT was already a common structure for disparate data formats, we were better able to identify sites with multiple allocations and, perhaps more importantly, allocations with multiple sites. This is particularly helpful when an allocation has sites which cross WMA boundaries, instead of just assigning the full water amount to a single WMA we are now able to divide the amount of water between the number of relevant WMAs. As it is now possible to identify allocations with water used in multiple WMAs, it is no longer practical to store this information within a single column. Instead the stAllocationToWMATab.csv file was created, which is an allocation by WMA matrix containing the percent Place of Use area overlap with each WMA. We then use this percentage to divide the allocation's flow amount between the given WMAs during the cumulation process to hopefully provide more realistic totals of water use in each area. However, not every state provides areas of water use, so like HarDWR v1.0, a hierarchical decision tree was used to assign each allocation to a WMA. First, if a WMA could be identified based on the allocation ID, then that WMA was used; typically, when available, this applied to the entire state and no further steps were needed. Second was the spatial analysis of Place of Use to WMAs. Third was a spatial analysis of the POD locations to WMAs, with the assumption that allocation's POD is within the WMA it should belong to; if an allocation still had multiple WMAs based on its POD locations, then the allocation's flow amount would be divided equally between all WMAs. The fourth, and final, process was to include water allocations which spatially fell outside of the state WMA boundaries. This could be due to several reasons, such as coordinate errors / imprecision in the POD location, imprecision in the WMA boundaries, or rights attached with features, such as a reservoir, which crosses state boundaries. To include these records, we decided for any POD which was within one kilometer of the state's edge would be assigned to the nearest WMA. Other Changes WestDAAT has Allowed In addition to a more nuanced and consistent method of assigning water right's data to WMAs, there are other benefits gained from using the WestDAAT dataset. Among those is a consistent categorization of a water right's legal status. In HarDWR v1.0, legal status was effectively ignored, which led to many valid concerns about the quality of the database related to the amounts of water the rights allowed to be claimed. The main issue was that rights with legal status' such as "application withdrawn", "non-active", or "cancelled" were included within HarDWR v1.0. These, and other water rights status' which were deemed to not be in use have been removed from this version of the database. Another major change has been the addition of the "unspecified water source category. This is water that can come from either surface water or groundwater, or the source of which is unknown. The addition of this source category brings the total number of categories to three. Due to reviewer feedback, we decided to add the acre-foot (AF) column so that the data may be more applicable to a wider audience. We added the ownerClassification column so that the data may be more applicable to a wider audience. File Descriptions The dataset is a series of various files organized by state sub-directories. In addition, each file begins with the state's name, in case the file is separate from its sub-directory for some reason. After the state name is the text which describes the contents of the file. Here is each file described in detail. Note that st is a placeholder for the state's name. stFullRecords_HarmonizedRights.csv: A file of the complete water records for each state. The column headers for each of this type of file are: state - The name of the state to which the allocations belong to. FIPS - The two digit numeric state ID code. siteID - The site location ID for POD locations. A site may have multiple allocations, which are the actual amount of water which can be drawn. In a simplified hypothetical, a farm stead may have an allocation for "irrigation" and an allocation for "domestic" water use, but the water is drawn from the same pumping equipment. It should be noted that many of the site ID appear to have been added by WaDE, and therefore may not be recognized by a given state's water rights database. allocationID - The allocation ID for the water right. For most states this is the water right ID, and what is recommended to use should a right be looked up on a given state's water rights database. The water amounts associated with these IDs tend to be finer scaled than those associated with siteID. It should be noted that some allocations may be extracted from multiple sites, particularly for larger Places of Use. ownerClassification - A classification of the types of owners for water rights. The most common is `Private` which incorporates a wide range of entities. Several classifications would be grouped into a government category, most of which are for the U.S. Federal Government. These allocations could be listed as "Federal", "United States of America", or as the names of any number of federal agencies. The last major grouping of entities is for "Native American"s. priorityDate - The date we use as the water right priority date for our modeling analysis. This is the legal priority date when it is available. However, for some rights, specifically from California and New Mexico, we used a pseudo priority date (e.g. well completion date or start of well drilling date) when a legal priority date was not available. The most questionable dates come from New Mexico, where the only date associated with certain water right records was the date the allocation was recorded in the database. As the allocation record creation tended to be within a few months of the filing of the application of the water right, from manually double checking the water rights, and our analysis focuses on aggregating water rights on the timescale of years, we determined it was acceptable to use such dates to include as many records as possible. primaryBeneficialUse - From the numerous state water use categories, WaDE categorized them into 21 categories WestDAAT. This column is the original WaDE category for the primary water use at the PoD site. allocationBeneficialUse - From the numerous state water use categories, WaDE categorized them into 21 categories for WestDAAT. This column is the original WaDE category

Economics↗

Integrating Data From In Vitro New Approach Methodologies for Developmental Neurotoxicity

Abstract In vivo developmental neurotoxicity (DNT) testing is resource intensive and lacks information on cellular processes affected by chemicals. To address this, DNT new approach methodologies (NAMs) are being evaluated, including: the microelectrode array neuronal network formation assay; and high-content imaging to evaluate proliferation, apoptosis, neurite outgrowth, and synaptogenesis. This work addresses 3 hypotheses: (1) a broad screening battery provides a sensitive marker of DNT bioactivity; (2) selective bioactivity (occurring at noncytotoxic concentrations) may indicate functional processes disrupted; and, (3) a subset of endpoints may optimally classify chemicals with in vivo evidence for DNT. The dataset was comprised of 92 chemicals screened in all 57 assay endpoints sourced from publicly available data, including a set of DNT NAM evaluation chemicals with putative positives (53) and negatives (13). The DNT NAM battery provides a sensitive marker of DNT bioactivity, particularly in cytotoxicity and network connectivity parameters. Hierarchical clustering suggested potency (including cytotoxicity) was important for classifying positive chemicals with high sensitivity (93%) but failed to distinguish patterns of disrupted functional processes. In contrast, clustering of selective values revealed informative patterns of differential activity but demonstrated lower sensitivity (74%). The false negatives were associated with several limitations, such as the maximal concentration tested or gaps in the biology captured by the current battery. This work demonstrates that this multi-dimensional assay suite provides a sensitive biomarker for DNT bioactivity, with selective activity providing possible insight into specific functional processes affected by chemical exposure and a basis for further research.

Carstens, Kelly E.↗

Extremely massive disc galaxies in the nearby Universe form through gas-rich minor mergers

ABSTRACT In our hierarchical structure-formation paradigm, the observed morphological evolution of massive galaxies – from rotationally supported discs to dispersion-dominated spheroids – is largely explained via galaxy merging. However, since mergers are likely to destroy discs, and the most massive galaxies have the richest merger histories, it is surprising that any discs exist at all at the highest stellar masses. Recent theoretical work by our group has used a cosmological, hydrodynamical simulation to suggest that extremely massive (M* > 1011.4 M⊙) discs form primarily via minor mergers between spheroids and gas-rich satellites, which create new rotational stellar components and leave discs as remnants. Here, we use UV-optical and H i data of massive galaxies, from the Sloan Digital Sky Survey, Galaxy Evolution Explorer, Dark Energy Camera Legacy Survey (DECaLS), and Arecibo Legacy Fast ALFA surveys, to test these theoretical predictions. Observed massive discs account for ∼13 per cent of massive galaxies, in good agreement with theory (∼11 per cent). ∼64 per cent of the observed massive discs exhibit tidal features, which are likely to indicate recent minor mergers, in the deep DECaLS images (compared to ∼60 per cent in their simulated counterparts). The incidence of these features is at least four times higher than in low-mass discs, suggesting that, as predicted, minor mergers play a significant (and outsized) role in the formation of these systems. The empirical star formation rates agree well with theoretical predictions and, for a small galaxy sample with H i detections, the H i masses and fractions are consistent with the range predicted by the simulation. The good agreement between theory and observations indicates that extremely massive discs are indeed remnants of recent minor mergers between spheroids and gas-rich satellites.

79 ASTRONOMY AND ASTROPHYSICS↗

Atomistic Simulation of Glasses and Amorphous Materials: Challenges and Opportunities for the Next Decade

Atomistic simulations have become indispensable tools for understanding glass structure, dynamics, and properties, yet persistent challenges limit their predictive power. This perspective examines three interconnected issues, namely glass formation procedures, interatomic potential development, and machine learning applications, which emerged from the 5th International Workshop on Challenges of Atomistic Simulations of Glasses and Amorphous Materials. We identify convergent community priorities for (i) standardized validation protocols, (ii) curated benchmark datasets with complete metadata, and (iii) open repositories for glasses. A systematic was forward is provided by a hierarchical validation framework for assessing the structural fidelity, property prediction, and behavioral realism of simulation techniques. Looking ahead, transformative advances are promised by the fusion of classical techniques with machine learning based approaches, for instance, by integrating swap Monte Carlo with machine-learning (ML) potentials, leveraging foundation models through transfer learning, and finetuning ML potentials with experimental data. Progress depends on the community committing to validated models, reproducible protocols, and sustained data sharing.

Krishnan, N. M. Anoop↗

Galaxy Cruise: Deep Insights into Interacting Galaxies in the Local Universe

Abstract We present the first results from GALAXY CRUISE, a community (or citizen) science project based on data from the Hyper Suprime-Cam Subaru Strategic Program (HSC-SSP). The current paradigm of galaxy evolution suggests that galaxies grow hierarchically via mergers, but our observational understanding of the role of mergers is still limited. The data from HSC-SSP are ideally suited to improve our understanding with improved identifications of interacting galaxies thanks to the superb depth and image quality of HSC-SSP. We launched a community science project, GALAXY CRUISE, in 2019 and have collected over two million independent classifications of 20686 galaxies at z < 0.2. We first characterize the accuracy of the participants’ classifications and demonstrate that it surpasses previous studies based on shallower imaging data. We then investigate various aspects of interacting galaxies in detail. We show that there is a clear sign of enhanced activities of super-massive black holes and star formation in interacting galaxies compared to those in isolated galaxies. The enhancement seems particularly strong for galaxies undergoing violent mergers. We also show that the mass growth rate inferred from our results is roughly consistent with the observed evolution of the stellar mass function. The second season of GALAXY CRUISE is currently underway and we conclude with future prospects. We make the morphological classification catalog used in this paper publicly available at the GALAXY CRUISE website, which will be particularly useful for machine-learning applications.

Tanaka, Masayuki↗

FilDReaMS: II. Application to the analysis of the relative orientations between filaments and the magnetic field in four Herschel fields

Context. Both simulations and observations of the interstellar medium show that the study of the relative orientations between filamentary structures and the magnetic field can bring new insight into the role played by magnetic fields in the formation and evolution of filaments and in the process of star formation. Aims. We provide a first application of FilDReaMS, the new method presented in the companion paper to detect and analyze filaments in a given image. The method relies on a template that has the shape of a rectangular bar with variable width. Our goal is to investigate the relative orientations between the detected filaments and the magnetic field. Methods. We apply FilDReaMS to a small sample of four Herschel fields (G210, G300, G82, G202) characterized by different Galactic environments and different evolutionary stages. First, we look for the most prevalent bar widths, and we examine the networks formed by filaments of different bar widths as well as their hierarchical organization. Second, we compare the filament orientations to the magnetic field orientation inferred from Planck polarization data and, for the first time, we study the statistics of the relative orientation angle as functions of both spatial scale and H2 column density. Results. We find preferential relative orientations in the four Herschel fields: small filaments with low column densities tend to be slightly more parallel than perpendicular to the magnetic field; in contrast, large filaments, which all have higher column densities, are oriented nearly perpendicular (or, in the case of G202, more nearly parallel) to the magnetic field. In the two nearby fields (G210 and G300), we observe a transition from mostly parallel to mostly perpendicular relative orientations at an H 2 column density ≃ 1.1 × 10 21 cm -2 and 1.4 × 10 21 cm -2 , respectively, consistent with the results of previous studies. Conclusions. Our results confirm the existence of a coupling between magnetic fields at cloud scales and filaments at smaller scale. They also illustrate the potential of combining Herschel and Planck observations, and they call for further statistical analyses with our dedicated method.

79 ASTRONOMY AND ASTROPHYSICS↗

Dark Energy Survey Year 3 results: optimized $w$CDM simulation-based inference with weak lensing map-level hybrid statistics

We present cosmological constraints from the Dark Energy Survey Year 3 (DES Y3) weak lensing data using hierarchical hybrid statistics within a Bayesian simulation-based inference framework that is based on the Gower Street simulations. To maximize the precision of the inference, we have developed a new, information-theory based, data compression of the weak lensing maps to just seven highly informative summary statistics. The hybrid scheme exploits the high information content of the power spectrum, compressing both the power spectrum and neural-based summaries that are designed to extract further information. Our simulation-based approach enables principled forward modelling of all major sources of systematic uncertainty and survey properties into realistic mock observations, including the survey mask, photometric redshift uncertainties, intrinsic galaxy alignments, multiplicative shear calibration bias, source galaxy clustering, non-Gaussian shape noise, and non-linear structure formation. The summary statistics are then used in a Bayesian simulation-based inference pipeline. The inference is validated through coverage tests and checks for robustness against baryonic feedback. Assuming a $w$CDM cosmology, our analysis yields $S_8 = 0.808 \pm 0.017$, $Ω_{\rm m} = 0.325 \pm 0.024$, and $w < -0.766$ (marginalized posterior 68 per cent credible intervals). This rigorous combination of information theory, physics- and neural network-based extreme data compression, and principled Bayesian analysis improves the figure of merit for $(Ω_{\rm m}, S_8, w)$ by 60 per cent over the previous state-of-the-art, and by almost a factor of 3 over two-point analyses of the same data. They are the most precise joint constraints on $(Ω_{\rm m}, S_8, w)$ from weak gravitational lensing data alone of any survey to date. We intend to apply this analysis to the more recent DES Y6 data.

Williamson, J. [University Coll. London]↗

Meteorological and hydrological parameters for 17 locations of meteorological stations of the East River Watershed

The Data Package includes a set of csv files of input meteorological parameters for locations of 17 meteorological stations within the East River watershed. These meteorological datasets were downloaded from the (1) PRISM database--monthly precipitation, air temperature (minimum, mean, and maximum), vapor pressure deficit (minimum and maximum), and dewpoint temperature, and (2) NCEP/NCAR Reanalysis database – wind database. The datasets were used for calculations of the Potential Evapotranspiration (ETo), Actual Evapotranspiration (ET), Standard Precipitation Index (SPI) , Standard Evapotranspiration-Precipitation Index (SPEI) for the period from 1966 to 2021. The main research questions addressed are: the evaluation of the long-term temporal trends of climatic parameters, hierarchical clustering, and areal mapping/zonation of the East River watershed. Calculations were conducted in the Rstudio environment. The dataset additionally includes a file-level metadata (flmd.csv) file that lists each file contained in the dataset with associated metadata; and a data dictionary (dd.csv) file that contains column/row headers used throughout the files along with a definition, units, and data type.The input datasets were downloaded from (a) PRISM database (the Northwest Alliance for Computational Science and Engineering at the Oregon State University), and (b) NCEP/NCAR Reanalysis database.

54 ENVIRONMENTAL SCIENCES↗