Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “pattern clustering”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 199 records · Page 11

Shear Instabilities as a Probe of Jupiter's Atmosphere

Linear wave patterns in Jupiter clouds with wavelengths strongly clustered around 300 km are commonly observed in the planet's equatorial atmosphere. We propose that the preferred wavelength is related to the thickness of an unstable shear layer within the clouds. We numerically analyze the linear stability of wavelike disturbances that have nonzero horizontal phase speeds in Jupiter's atmosphere and find that. if the static stability in the strongly clustered around 300 km are commonly observed in the planet's equatorial atmosphere. We propose that the preferred wavelength is related to the thickness of an unstable shear layer within the clouds. We numerically analyze the linear stability of wavelike disturbances that have nonzero horizontal phase speeds in Jupiter's atmosphere and find that. if the static stability in the shear layer is very low (but still nonnegative), a deep vertical shear layer like the one measured by the Galileo probe can generate the instabilities. The fastest growing waves grow exponentially within an hour, and their wavelengths match the observations. Close to zero values of static stability that permit the growth of instabilities are within the range of values measured by the Galileo probe in a hot spot. Our model probes Jupiter's equatorial atmosphere below the cloud deck and suggests that thick regions of wind shear and low static stability exist outside hot spots.

Bosak, Tanja↗

Double diffraction imaging of x-ray induced structural dynamics in single free nanoparticles

Abstract Because of their high photon flux, x-ray free-electron lasers (FEL) allow to resolve the structure of individual nanoparticles via coherent diffractive imaging (CDI) within a single x-ray pulse. Since the inevitable rapid destruction of the sample limits the achievable resolution, a thorough understanding of the spatiotemporal evolution of matter on the nanoscale following the irradiation is crucial. We present a technique to track x-ray induced structural changes in time and space by recording two consecutive diffraction patterns of the same single, free-flying nanoparticle, acquired separately on two large-area detectors opposite to each other, thus examining both the initial and evolved particle structure. We demonstrate the method at the extreme ultraviolet (XUV) and soft x-ray Free-electron LASer in Hamburg (FLASH), investigating xenon clusters as model systems. By splitting a single XUV pulse, two diffraction patterns from the same particle can be obtained. For focus intensities of about 2 × 10 12 W cm −2 we observe still largely intact clusters even at the longest delays of up to 650 picoseconds of the second pulse, indicating that in the highly absorbing systems the damage remains confined to one side of the cluster. Instead, in case of five times higher flux, the diffraction patterns show clear signatures of disintegration, namely increased diameters and density fluctuations in the fragmenting clusters. Future improvements to the accessible range of dynamics and time resolution of the approach are discussed.

Sauppe, M. (ORCID:0000000163123102)↗

Evidence for C and Mg variations in the GD-1 stellar stream

ABSTRACT Dynamically cold stellar streams are the relics left over from globular cluster dissolution. These relics offer a unique insight into a now fully disrupted population of ancient clusters in our Galaxy. Using a combination of Gaia eDR3 proper motions, optical and near-UV colours, we select a sample of likely Red Giant Branch stars from the GD-1 stream for medium-low resolution spectroscopic follow-up. Based on radial velocity and metallicity, we are able to find 14 new members of GD-1, 5 of which are associated with the spur and blob/cocoon off-stream features. We measured C-abundances to probe for abundance variations known to exist in globular clusters. These variations are expected to manifest in a subtle way in globular clusters with such low masses ($\sim 10^4\,{\rm ~\textrm {M}_\odot }$) and metallicities ([Fe/H] ∼ −2.1 dex). We find that the C-abundances of the stars in our sample display a small but significant (3σ level) spread. Furthermore, we find ∼3σ variation in Mg-abundances among the stars in our sample that have been observed by APOGEE. These abundance patterns match the ones found in Galactic globular clusters of similar metallicity. Our results suggest that GD-1 represents another fully disrupted low-mass globular cluster where light-element abundance spreads have been found.

79 ASTRONOMY AND ASTROPHYSICS↗

Comprehensive Genetic Characterization of Four Novel HIV-1 Circulating Recombinant Forms (CRF129_56G, CRF130_A1B, CRF131_A1B, and CRF138_cpx): Insights from Molecular Epidemiology in Cyprus

Molecular investigations of the HIV-1 pol region (2253–5250 in the HXB2 genome) were conducted on sequences obtained from 331 individuals infected with HIV-1 in Cyprus between 2017 and 2021. This study unveiled four distinct HIV-1 putative transmission clusters, encompassing 19 previously unidentified HIV-1 recombinants. These recombinants, each comprising eight, three, four, and four sequences, respectively, did not align with previously established Circulating Recombinant Forms (CRFs). To characterize these novel HIV-1 recombinants, near-full-length genome sequences were successfully obtained for 16 of the 19 recombinants (790–8795 in the HXB2 genome) using an in-house-developed RT-PCR assay. Phylogenetic analyses, employing MEGAX and Cluster-Picker, along with confirmatory neighbor-joining tree analyses of subregions, were conducted to identify distinct clusters and determine subtypes. The uniqueness of the HIV-1 recombinants was evident in their exclusive clustering within generated maximum likelihood trees. Recombination analyses highlighted the distinct chimeric nature of these recombinants, with consistent mosaic patterns observed across all sequences within each of the four putative transmission clusters. Conclusive genetic characterization identified four novel HIV-1 CRFs: CRF129_56G, CRF130_A1B, CRF131_A1B, and CRF138_cpx. CRF129_56G exhibited two recombination breakpoints and three fragments of subtypes CRF56_cpx and G. Both CRF130_A1B and CRF131_A1B featured seven recombination breakpoints and eight fragments of subtypes A1 and B. CRF138_cpx displayed five recombination breakpoints and six fragments of subtypes CRF22_01A1 and F2, along with an unclassified fragment. Additional BLAST analyses identified a Unique Recombinant Form (URF) of CRF138_cpx with three additional recombination sites, involving subtype F2, a fragment of unknown subtype origin, and CRF138_cpx. Post-identification, all putative transmission clusters remained active, with CRF130_A1B, CRF131_A1B, and CRF138_cpx clusters exhibiting further growth. Furthermore, international connections were identified through BLAST analyses, linking one sequence from the USA to the CRF130_A1B strain, and three sequences from Belgium and Cameroon to the CRF138_cpx strain. This study contributes valuable insights into the dynamic landscape of HIV-1 diversity and transmission patterns, emphasizing the need for ongoing molecular surveillance and global collaboration in tracking emerging viral variants.

60 APPLIED LIFE SCIENCES↗

Microhydration of the metastable N -Protomer of 4-Aminobenzoic acid by condensation at 80 K: H/D exchange without conversion to the more stable O- protomer

4-Aminobenzoic acid (4ABA) is a model scaffold for studying solvent-mediated proton transfer. Although protonation at the carboxylic group (O-protomer) is energetically favored in the gas phase, the N-protomer, where the proton remains on the amino group, can be kinetically trapped by electrospray ionization of 4ABA in an aprotic solvent such as acetonitrile. Here we report the formation of the hydrated deuterium isotopologues of the N-protomers, RND 3 + ·(H 2 O) n=1-3 , (R=C 6 H 4 COOD), which are generated by condensing water molecules onto the bare N-protomers in a liquid nitrogen cooled, radiofrequency octopole ion trap at 80 K. The product clusters are then transferred to a 20 K cryogenic ion trap where they are tagged with weakly bound D 2 molecules. The structures of these clusters are determined by analysis of their vibrational patterns obtained by resonant IR photodissociation. The resulting patterns confirm that the metastable N-protomer configuration remains intact even when warmed by sequential condensation of water molecules. Attachment of H 2 O molecules onto the RND 3 + head group also affords the opportunity to explore the possibility of H/D exchange between the acid scaffold and the proximal water network. The spectroscopic results establish that although the RND 3 + ·(H 2 O) n=1,2 clusters are formed without H/D exchange, the n = 3 cluster exhibits about 10% H/D exchange as evidenced by the appearance of the telltale HOD bands. Furthermore, the site of exchange on the acid is determined to be the acidic OH by the emergence of the OH stretching fundamental in the -COOH motif.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Iron-sulfur clusters: the road to room temperature

Abstract Iron-sulfur proteins perform a wide variety of reactions central to the metabolisms of all living organisms. Foundational to their reaction chemistry are the rich electronic structures of their constituent Fe-S clusters, which differ in important ways from the active sites of mononuclear Fe enzymes. In this perspective, we summarize the essential electronic structure features that make Fe-S clusters unique, and point to the need for studies aimed at understanding the electronic basis for their reactivity under physiological conditions. Specifically, at ambient temperature, both the ground state and a large number of excited states are thermally populated, and thus a complete understanding of Fe-S cluster reactivity must take into account the properties, energies, and reactivity patterns of these excited states. We highlight prior research toward characterizing the low-energy excited states of Fe-S clusters that has established what is now a consensus model of these excited state manifolds and the bonding interactions that give rise to them. In particular, we discuss the low-energy alternate spin states and valence electron configurations that occur in Fe-S clusters of varying nuclearities, and finally suggest that there may be unrecognized functional roles for these states. Graphical abstract

Skeel, Brighton A. (ORCID:000000018458088X)↗

Bioinformatics of cyanophycin metabolism genes and characterization of promiscuous isoaspartyl dipeptidases that catalyze the final step of cyanophycin degradation

Cyanophycin is a bacterial biopolymer used for storage of fixed nitrogen. It is composed of a backbone of L-aspartate residues with L-arginines attached to each of their side chains. Cyanophycin is produced by cyanophycin synthetase 1 (CphA1) using Arg, Asp and ATP, and is degraded in two steps. First, cyanophycinase breaks down the backbone peptide bonds, releasing β-Asp-Arg dipeptides. Then, these dipeptides are broken down into free Asp and Arg by enzymes with isoaspartyl dipeptidase activity. Two bacterial enzymes are known to possess promiscuous isoaspartyl dipeptidase activity: isoaspartyl dipeptidase (IadA) and isoaspartyl aminopeptidase (IaaA). We performed a bioinformatic analysis to investigate whether genes for cyanophycin metabolism enzymes cluster together or are spread around the microbial genomes. Many genomes showed incomplete contingents of known cyanophycin metabolizing genes, with different patterns in various bacterial clades. Cyanophycin synthetase and cyanophycinase are usually clustered together when recognizable genes for each are found within a genome. Cyanophycinase and isoaspartyl dipeptidase genes typically cluster within genomes lacking cphA1. About one-third of genomes with genes for CphA1, cyanophycinase and IaaA show these genes clustered together, while the proportion is around one-sixth for CphA1, cyanophycinase and IadA. We used X-ray crystallography and biochemical studies to characterize an IadA and an IaaA from two such clusters, in Leucothrix mucor and Roseivivax halodurans, respectively. The enzymes retained their promiscuous nature, showing that being associated with cyanophycin-related genes did not make them specific for β-Asp-Arg dipeptides derived from cyanophycin degradation.

59 BASIC BIOLOGICAL SCIENCES↗

The many-body expansion for aqueous systems revisited: III. Hofmeister ion – water interactions

We report a Many Body Energy (MBE) analysis of aqueous ionic clusters containing anions and cations at the two opposite ends of the Hofmeister series, viz. the kosmotropes Ca2+, SO42- and chaotropes NH4+ and ClO4- with 9 water molecules to quantify the how these ions in altering the interaction between the water molecules in their immediate surrounding. The current results are contrasted to the ones reported earlier for water clusters as well as for alkali metal and halide ion aqueous clusters of the same size, which lie in the middle of the Hofmeister series. Through this analysis, noteworthy differences between the MBE of kosmotropes and chaotropes were identified. The MBE of kosmotropes is dominated by ion-water interactions that extends beyond the 4-body term, the point at which the MBE of pure water converges. The percentage contribution of the 2- B to the total cluster binding energy is noticeably larger. The disruption due to the dominant ion results in weak, unfavorable water-water interactions. The MBE for chaotropes, on the other hand, was found to converge more quickly as it more closely resembles that of pure water clusters. Chaotropes exhibit weaker overall binding energies and ion-water interactions with more favorable water-water interactions, somewhat recovering the pattern of the 2-4 body terms exemplified by pure water clusters. More importantly, both kosmotropic and chaotropic ions exhibit an anticorrelation between the 2-B ion-water (I-W) and water-water (W-W) interactions as well as between the 3-B (I-W-W) and (I-W) interactions. The consideration of two different structural arrangements (ion inside and outside of a water cluster) suggests that fully solvated (ion inside) chaotropes disrupt the hydrogen bonding network in a similar manner as partially solvated (ion outside) kosmotropes and offer useful insights into the modeling requirements of bulk vs. an interface. Finally, the 2-B contribution to the total Basis Set Superposition Error (BSSE) correction for the kosmotropic and chaotropic ions follows the previously reported erf profile vs. intermolecular distance. When scaled for the corresponding dimer energies and distances, a single profile fits the current results together with all previously reported ones for the pure water and halide water clusters.

Herman, Kristina M.↗

Fractal dimension of bioconvection patterns

Shallow cultures of the motile algal strain, Euglena gracilis, were concentrated to 2 x 10 to the 6th organisms per ml and placed in constant temperature water baths at 24 and 38 C. Bioconvective patterns formed an open two-dimensional structure with random branches, similar to clusters encountered in the diffusion-limited aggregation (DLA) model. When averaged over several example cultures, the pattern was found to have no natural length scale, self-similar branching, and a fractal dimension (d about 1.7). These agree well with the two-dimensional DLA.

Noever, David A.↗

Value, Cost, and Sharing: Open Issues in Constrained Clustering

Clustering is an important tool for data mining, since it can identify major patterns or trends without any supervision (labeled data). Over the past five years, semi-supervised (constrained) clustering methods have become very popular. These methods began with incorporating pairwise constraints and have developed into more general methods that can learn appropriate distance metrics. However, several important open questions have arisen about which constraints are most useful, how they can be actively acquired, and when and how they should be propagated to neighboring points. This position paper describes these open questions and suggests future directions for constrained clustering research.

constraints↗

Improve Data Mining and Knowledge Discovery Through the Use of MatLab

Data mining is widely used to mine business, engineering, and scientific data. Data mining uses pattern based queries, searches, or other analyses of one or more electronic databases/datasets in order to discover or locate a predictive pattern or anomaly indicative of system failure, criminal or terrorist activity, etc. There are various algorithms, techniques and methods used to mine data; including neural networks, genetic algorithms, decision trees, nearest neighbor method, rule induction association analysis, slice and dice, segmentation, and clustering. These algorithms, techniques and methods used to detect patterns in a dataset, have been used in the development of numerous open source and commercially available products and technology for data mining. Data mining is best realized when latent information in a large quantity of data stored is discovered. No one technique solves all data mining problems; challenges are to select algorithms or methods appropriate to strengthen data/text mining and trending within given datasets. In recent years, throughout industry, academia and government agencies, thousands of data systems have been designed and tailored to serve specific engineering and business needs. Many of these systems use databases with relational algebra and structured query language to categorize and retrieve data. In these systems, data analyses are limited and require prior explicit knowledge of metadata and database relations; lacking exploratory data mining and discoveries of latent information. This presentation introduces MatLab(R) (MATrix LABoratory), an engineering and scientific data analyses tool to perform data mining. MatLab was originally intended to perform purely numerical calculations (a glorified calculator). Now, in addition to having hundreds of mathematical functions, it is a programming language with hundreds built in standard functions and numerous available toolboxes. MatLab's ease of data processing, visualization and its enormous availability of built in functionalities and toolboxes make it suitable to perform numerical computations and simulations as well as a data mining tool. Engineers and scientists can take advantage of the readily available functions/toolboxes to gain wider insight in their perspective data mining experiments.

Shaykhian, Gholam Ali↗

Improve Data Mining and Knowledge Discovery through the use of MatLab

Data mining is widely used to mine business, engineering, and scientific data. Data mining uses pattern based queries, searches, or other analyses of one or more electronic databases/datasets in order to discover or locate a predictive pattern or anomaly indicative of system failure, criminal or terrorist activity, etc. There are various algorithms, techniques and methods used to mine data; including neural networks, genetic algorithms, decision trees, nearest neighbor method, rule induction association analysis, slice and dice, segmentation, and clustering. These algorithms, techniques and methods used to detect patterns in a dataset, have been used in the development of numerous open source and commercially available products and technology for data mining. Data mining is best realized when latent information in a large quantity of data stored is discovered. No one technique solves all data mining problems; challenges are to select algorithms or methods appropriate to strengthen data/text mining and trending within given datasets. In recent years, throughout industry, academia and government agencies, thousands of data systems have been designed and tailored to serve specific engineering and business needs. Many of these systems use databases with relational algebra and structured query language to categorize and retrieve data. In these systems, data analyses are limited and require prior explicit knowledge of metadata and database relations; lacking exploratory data mining and discoveries of latent information. This presentation introduces MatLab(TradeMark)(MATrix LABoratory), an engineering and scientific data analyses tool to perform data mining. MatLab was originally intended to perform purely numerical calculations (a glorified calculator). Now, in addition to having hundreds of mathematical functions, it is a programming language with hundreds built in standard functions and numerous available toolboxes. MatLab's ease of data processing, visualization and its enormous availability of built in functionalities and toolboxes make it suitable to perform numerical computations and simulations as well as a data mining tool. Engineers and scientists can take advantage of the readily available functions/toolboxes to gain wider insight in their perspective data mining experiments.

Shaykahian, Gholan Ali↗

Sequence Length of HIV-1 Subtype B Increases over Time: Analysis of a Cohort of Patients with Hemophilia over 30 Years

We aimed to investigate whether the sequence length of HIV-1 increases over time. We performed a longitudinal analysis of full-length coding region sequences (FLs) during an HIV-1 outbreak among patients with hemophilia and local controls infected with the Korean subclade B of HIV-1 (KSB). Genes were amplified by overlapping RT-PCR or nested PCR and subjected to direct sequencing. Overall, 141 FLs were sequentially determined over 30 years in 62 KSB-infected patients. Phylogenetic analysis indicated that within KSB, two FLs from plasma donors O and P comprised two clusters, together with 8 and 12 patients with hemophilia, respectively. Signature pattern analysis of the KSB of HIV-1 revealed 91 signature nucleotide residues (1.1%). In total, 48 and 43 signature nucleotides originated from clusters O and P, respectively. Six positions contained 100% specific nucleotide(s) in clusters O and P. In-depth FL analysis for over 30 years indicated that the KSB FL significantly increased over time before combination antiretroviral therapy (cART) and decreased with cART. This increase occurred due to the significant increase in env and nef genes, originating in the variable regions of both genes. The increase in sequence length of HIV-1 over time suggests an evolutionary direction.

59 BASIC BIOLOGICAL SCIENCES↗

Predictions of a population of cataclysmic variables in globular clusters

We have studied the number of cataclysmic variables (CVs) that should be active in globular clusters during the present epoch as a result of binary formation via two-body tidal capture. We predict the orbital period and luminosity distributions of CVs in globular clusters. The results arebased on Monte Carlo simulations combined with evolution calculations appropriate to each system formed during the lifetime of two specific globular clusters, omega Cen and 47 Tuc. From our study of these two clusters, which represent the range of core densities and states of mass segregation that are likely to be interesting, we extrapolate our results to the Galactic globlular cluster system. Although there is at present little direct observational evidence of CVs in globular clusters, we find that there should be a large number of active systems. We predict that there should be more than approximately 100 CVs in both 47 Tuc and omega Cen and several thousand in the Galactic globular cluster system. These numbers are based on two-body processes alone and represent a lower bound on the number of systems that may have been formed as a result of stellar interaction within globular clusters. The relation between these calculations and the paucity of optically detected CVs in globular clusters is discussed. Should future observations fail to find convincing evidence of a substantial population of cluster CVs, then the two-body tidal capture scenario is likely to be seriously constrained. Of the CVs we espect in 47 Tuc and omega Cen, approximately 45 and 20, respectively, should have accretion luminosities above 10(exp 33) ergs/s. If one utilizes a relation for converting accretion luminosity to hard X-ray luminosity that is based on observations of Galactic plane CVs, even these sources will not exhibit X-ray luminosities above 10(exp 33) ergs/s. While we cannot account directly for the most luminous subset of the low-luminosity globular cluster X-ray sources without assuming an evolutionary pattern that is different from that of the majority of CVs in the disk, we are able to account for all of the observed lower luminosity subset of these sources, many of which have been recently discovered through ROSAT observations. In order for our predicted integrated cluster X-ray luminosities to be consistent with observational upper limits, the relation between accretion and X-ray luminosities should be something like that inferred from the Galactic plane population of CVs. Our calculations predict a large number of systems with L(sub acc) is less than 10(exp 32) ergs/s. Although our calculations imply that globular clusters should have an enhancement of CVs relative to the number thought to be present in the Galactic disk, this enhancement is at most roughly an order of magnitude, not comparable to the factor of approximately 100 for low-mass X-ray binaries (LMXBs).

Di Stefano, R.↗

Describing patterns of familial cancer risk in subfertile men using population pedigree data

STUDY QUESTION Can we simultaneously assess risk for multiple cancers to identify familial multicancer patterns in families of azoospermic and severely oligozoospermic men? SUMMARY ANSWER Here, distinct familial cancer patterns were observed in the azoospermia and severe oligozoospermia cohorts, suggesting heterogeneity in familial cancer risk by both type of subfertility and within subfertility type. WHAT IS KNOWN ALREADY Subfertile men and their relatives show increased risk for certain cancers including testicular, thyroid, and pediatric. STUDY DESIGN, SIZE, DURATION A retrospective cohort of subfertile men (N = 786) was identified and matched to fertile population controls (N = 5674). Family members out to third-degree relatives were identified for both subfertile men and fertile population controls (N = 337 754). The study period was 1966–2017. Individuals were censored at death or loss to follow-up, loss to follow-up occurred if they left Utah during the study period. PARTICIPANTS/MATERIALS, SETTING, METHODS Azoospermic (0 × 10 6 /mL) and severely oligozoospermic (<1.5 × 10 6 /mL) men were identified in the Subfertility Health and Assisted Reproduction and the Environment cohort (SHARE). Subfertile men were age- and sex-matched 5:1 to fertile population controls and family members out to third-degree relatives were identified using the Utah Population Database (UPDB). Cancer diagnoses were identified through the Utah Cancer Registry. Families containing ≥10 members with ≥1 year of follow-up 1966–2017 were included (azoospermic: N = 426 families, 21 361 individuals; oligozoospermic: N = 360 families, 18 818 individuals). Unsupervised clustering based on standardized incidence ratios for 34 cancer phenotypes in the families was used to identify familial multicancer patterns; azoospermia and severe oligospermia families were assessed separately. MAIN RESULTS AND THE ROLE OF CHANCE Compared to control families, significant increases in cancer risks were observed in the azoospermia cohort for five cancer types: bone and joint cancers hazard ratio (HR) = 2.56 (95% CI = 1.48–4.42), soft tissue cancers HR = 1.56 (95% CI = 1.01–2.39), uterine cancers HR = 1.27 (95% CI = 1.03–1.56), Hodgkin lymphomas HR = 1.60 (95% CI = 1.07–2.39), and thyroid cancer HR = 1.54 (95% CI = 1.21–1.97). Among severe oligozoospermia families, increased risk was seen for three cancer types: colon cancer HR = 1.16 (95% CI = 1.01–1.32), bone and joint cancers HR = 2.43 (95% CI = 1.30–4.54), and testis cancer HR = 2.34 (95% CI = 1.60–3.42) along with a significant decrease in esophageal cancer risk HR = 0.39 (95% CI = 0.16–0.97). Thirteen clusters of familial multicancer patterns were identified in families of azoospermic men, 66% of families in the azoospermia cohort showed population-level cancer risks, however, the remaining 12 clusters showed elevated risk for 2-7 cancer types. Several of the clusters with elevated cancer risks also showed increased odds of cancer diagnoses at young ages with six clusters showing increased odds of adolescent and young adult (AYA) diagnosis [odds ratio (OR) = 1.96–2.88] and two clusters showing increased odds of pediatric cancer diagnosis (OR = 3.64–12.63). Within the severe oligozoospermia cohort, 12 distinct familial multicancer clusters were identified. All 12 clusters showed elevated risk for 1–3 cancer types. An increase in odds of cancer diagnoses at young ages was also seen in five of the severe oligozoospermia familial multicancer clusters, three clusters showed increased odds of AYA diagnosis (OR = 2.19–2.78) with an additional two clusters showing increased odds of a pediatric diagnosis (OR = 3.84–9.32). LIMITATIONS, REASONS FOR CAUTION Although this study has many strengths, including population data for family structure, cancer diagnoses and subfertility, there are limitations. First, semen measures are not available for the sample of fertile men. Second, there is no information on medical comorbidities or lifestyle risk factors such as smoking status, BMI, or environmental exposures. Third, all of the subfertile men included in this study were seen at a fertility clinic for evaluation. These men were therefore a subset of the overall population experiencing fertility problems and likely represent those with the socioeconomic means for evaluation by a physician. WIDER IMPLICATIONS OF THE FINDINGS This analysis leveraged unique population-level data resources, SHARE and the UPDB, to describe novel multicancer clusters among the families of azoospermic and severely oligozoospermic men. Distinct overall multicancer risk and familial multicancer patterns were observed in the azoospermia and severe oligozoospermia cohorts, suggesting heterogeneity in cancer risk by type of subfertility and within subfertility type. Describing families with similar cancer risk patterns provides a new avenue to increase homogeneity for focused gene discovery and environmental risk factor studies. Such discoveries will lead to more accurate risk predictions and improved counseling for patients and their families. STUDY FUNDING/COMPETING INTEREST(S) This work was funded by GEMS: Genomic approach to connecting Elevated germline Mutation rates with male infertility and Somatic health (Eunice Kennedy Shriver National Institute of Child Health and Human Development (NICHD): R01 HD106112). The authors have no conflicts of interest relevant to this work.

60 APPLIED LIFE SCIENCES↗

The superatomic state beyond conventional magic numbers: Ligated metal chalcogenide superatoms

The field of cluster science is drawing increasing attention due to the strong size and composition-dependent properties of clusters and the exciting prospect of clusters serving as the building blocks for materials with tailored properties. However, identifying a unifying central paradigm that provides a framework for classifying and understanding the diverse behaviors is an outstanding challenge. One such central paradigm is the superatom concept that was developed for metallic and ligand-protected metallic clusters. The periodic electronic and geometric closed shells in clusters result in their properties being based on the stability they gain when they achieve closed shells. This stabilization results in the clusters having a well-defined valence allowing them to be classified as superatoms – thus, extending the periodic table to a third dimension. This perspective focuses on extending the superatomic concept to ligated metal-chalcogen clusters that have recently been synthesized in solutions and form assemblies with counterions that have wide-ranging applications. Here we illustrate that the periodic patterns emerge in the electronic structure of ligated metal-chalcogenide clusters. The stabilization gained by the closing of their electronic shells allows for the prediction of their redox properties. Further investigations reveal how the selection of ligands may control the redox properties of the superatoms. These ligated clusters may serve as chemical dopants for two-dimensional semiconductors to control their transport characteristics. Superatomic molecules of multiple metal-chalcogen superatoms allow for the formation of nano pn junctions ideal for directed transport and photon harvesting. As a result, the perspective outlines future developments, including the synthesis of magnetic superatoms.

36 MATERIALS SCIENCE↗

Signature of a Massive Rotating Metal-poor Star Imprinted in the Phoenix Stellar Stream

The Phoenix stellar stream has a low intrinsic dispersion in velocity and metallicity that implies the progenitor was probably a low-mass globular cluster. In this work we use Magellan/Magellan Inamori Kyocera Echelle (MIKE) high-dispersion spectroscopy of eight Phoenix stream red giants to confirm this scenario. In particular, we find negligible intrinsic scatter in metallicity ($σ$([FE II/H]) = ${0.04}_{-0.03}^{+0.11}$) and a large peak-to-peak range in [Na/Fe] and [Al/Fe] abundance ratios, consistent with the light element abundance patterns seen in the most metal-poor globular clusters. However, unlike any other globular cluster, we also find an intrinsic spread in [Sr II/Fe] spanning ~1 dex, while [Ba II/Fe] shows nearly no intrinsic spread ($σ$([Ba II/H]) = ${0.03}_{-0.02}^{+0.10}$). This abundance signature is best interpreted as slow-neutron-capture element production from a massive fast-rotating metal-poor star (15–20 $M$ ⊙ , $v$ ini /$v$ crit = 0.4, [Fe/H] = -3.8). The low inferred cluster mass suggests the system would have been unable to retain supernovae ejecta, implying that any massive fast-rotating metal-poor star that enriched the interstellar medium must have formed and evolved before the globular cluster formed. Furthermore, neutron-capture element production from asymptotic giant branch stars or magneto-rotational instabilities in core-collapse supernovae provide poor fits to the observations. We also report one Phoenix stream star to be a lithium-rich giant ($A$(Li) = 3.1 ± 0.1). At [Fe/H ] = -2.93; it is among the most metal-poor lithium-rich giants known.

79 ASTRONOMY AND ASTROPHYSICS↗

Incorporating space and time into random forest models for analyzing geospatial patterns of drug-related crime incidents in a major U.S. metropolitan area

The opioid crisis has hit American cities hard, and research on spatial and temporal patterns of drug-related activities including detecting and predicting clusters of crime incidents involving particular types of drugs is useful for distinguishing hot zones where drugs are present that in turn can further provide a basis for assessing and providing related treatment services. In this study, we investigated spatiotemporal patterns of more than 52,000 reported incidents of drug-related crime at block group granularity in Chicago, IL between 2016 and 2019. We applied a space-time analysis framework and machine learning approaches to build a model using training data that identified whether certain locations and built environment and sociodemographic factors were correlated with drug-related crime incident patterns, and establish the top contributing factors that underlaid the trends. Space and time, together with multiple driving factors, were incorporated into a random forest model to analyze these changing patterns. We accommodated both spatial and temporal autocorrelation in the model learning process to assist with capturing the changes over time and tested the capabilities of the space-time random forest model by predicting drug-related activity hot zones. Overall, we focused particularly on crime incidents that involved heroin and synthetic drugs as these have been key drug types that have highly impacted cities during the opioid crisis in the U.S.

97 MATHEMATICS AND COMPUTING↗