Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “co-occurrence”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

HydroFish: freshwater fish co-occurrence with hydropower plants and non-powered dams in conterminous United States sub-basins

The HydroFish dataset lists all existing hydropower plants (EHAs) and non-powered dams (NPDs; ≥ 0.001 MW potential nominal capacity), delineates the hydrologic sub-basins in which they are situated, and then lists all freshwater fish species reported to occur in those sub-basins. This dataset was compiled using the HydroBio dataset (https://hydrosource.ornl.gov/data/datasets/hydrobio/) and contains 24 total variables that describe hydrologic sub-basins, each unique EHA (plant ID and name, geographic coordinates, permit type, capacity, etc.) and NPD (ID value, known names, geographic coordinates, and estimated potential nominal capacity), and freshwater fish species in the sub-basin (common and scientific name, origin, and migratory and threat status). The HydroBio dataset was built using Oak Ridge National Laboratory’s Existing Hydropower Assets Dataset (2024 version) and Non-powered Dam Technical Potential Dataset (2024 version), and NatureServe’s fish species distribution dataset (2023 version). The dataset also contains summary variables that report the unique number of EHAs, NPDs, and freshwater fish species per sub-basin. The HydroFish dataset contains two unique data files: 1) a .csv metadata file describing the dataset variables, and 2) a .csv data file containing the actual dataset. Note that there may be many rows per unique existing hydropower plant or non-powered dam given that distinct species are listed per existing plant or NPD per sub-basin. The dataset is downloadable as a zip file containing the metadata and dataset files.

Bozeman, Bryan [Oak Ridge National Laboratory (ORN↗

Co-Occurring Atmospheric Features and Their Contributions to Precipitation Extremes

Object-based identification algorithms for atmospheric features are commonly utilized to attribute global precipitation. This study employs a systematic approach to examine feature co-occurrences and their relationships to mean and extreme precipitation. Four features are identified using existing data sets for atmospheric rivers (ARs), mesoscale convective systems (MCSs), low-pressure systems (LPSs), and fronts (FTs). Often, a single atmospheric phenomenon satisfies the criteria set by multiple feature identification algorithms, yielding an association between precipitation and multiple features. Over the extra-tropics, the number of features attributed to a single event typically increases with precipitation intensity. Over two-thirds of the precipitation is from co-occurring features, with a considerable fraction related to AR-FT co-occurrences. Over the tropics, about one-quarter of precipitation is associated with co-occurring features, with LPS-MCS co-occurrences contributing substantially in monsoon regions. MCSs are the leading single-feature contributors over tropical land and oceans. In the extra-tropics, FTs, ARs, and their co-occurrences account for over half of the total precipitation over oceans. AR-FT-MCS and FT-MCS co-occurrences contribute to extremes (precipitation exceeding the 95th percentile) over both oceans (over 30%) and land (over 20%). Any combination of features involving MCSs shows a larger contribution to high percentiles of precipitation intensity. A case analysis indicates that AR-FT-MCS co-occurrences exhibit convective instability and deep vertical motion, suggesting that the feature trackers and reanalysis are capturing physics relevant to both convective and frontal systems. The results here emphasize the need for simultaneous identifications of multiple features when attributing precipitation to atmospheric phenomena.

54 ENVIRONMENTAL SCIENCES↗

Compound Mesoscale Convective Systems and Low‐Pressure Systems in Tropical Monsoon Regions: Assessing Their Meteorology and Precipitation

Mesoscale convective systems (MCS) and low-pressure systems (LPS) are both strongly associated with precipitation across the regions where they occur, particularly within global monsoon systems; however, their co-occurrence and its relationship to precipitation have not been systematically examined. Here, we use LPS and MCS trackers to detect compound MCS and LPS events in five monsoon regions and assess the association of this co-occurrence with anomalies of winds, precipitation, and other atmospheric variables. Additionally, we investigate the spatial distribution of precipitating MCS and LPS events. Our results show that most (∼60%) MCS and LPS co-occurrences are located in the lower latitudes, where they contribute up to 40% of annual precipitation. We find that compound events generally produce more extreme precipitation than MCS-only or LPS-only events. Furthermore, our assessment of the synoptic and mesoscale composites reveals that the underlying dynamics of compound events exhibit anomalously positive convective available potential energy and an anomalously low-pressure within the location of the event. In terms of the synoptic environment of the features, the compound MCS and LPS events are associated with inverted troughs in three out of five monsoon locations assessed.

Environmental sciences↗

Quantitative stable isotope probing (qSIP) and cross-domain networks reveal bacterial-fungal interactions in the hyphosphere

Interactions between fungi and bacteria have the potential to substantially influence soil carbon dynamics in soil, but we have yet to fully identify these interactions and partners in their natural environment. In this study, we stacked two powerful methods, 13 C quantitative stable isotope probing (qSIP) and cross-domain co-occurrence network, to identify interacting fungi and bacteria in a California grassland soil. We used in-field whole plant 13 CO 2 labeling along with sand-filled ingrowth bags (that trap fungi and hyphae-associated bacteria) to amplify the signal of fungal-bacterial interactions, separate from the bulk soil background. We found a total of 54 bacterial ASVs and 9 fungal OTUs that were significantly 13 C-enriched. These were saprotrophic and biotrophic fungi, and motile, sometimes predatory bacteria. Among these, 70% of all 13 C-enriched bacteria identified were motile. Notably, we detected fungal-bacterial network links between a fungal OTU of the genus Alternaria and several bacterial ASVs of the genera Bacteriovorax, Mucilaginibacter, and Flavobacterium, providing empirical evidence of their direct interactions through C exchange. We observed a strong positive co-occurrence pattern between predatory bacteria of the phylum Bdellovibrionota and fungal OTUs, suggesting the transfer of C across the soil food web. To date, our ability to associate microbial co-occurrence network patterns with biological interactions is limited, but the incorporation of qSIP allowed us to more precisely detect interacting partners by narrowing in on the taxa that were actively incorporating plant-fixed, fungal-transported labeled substrates. Together, these approaches can help build a mechanistic understanding of the complex nature of fungal-bacterial interactions in soil.

59 BASIC BIOLOGICAL SCIENCES↗

Statistical relationships across epigenomes using large-scale hierarchical clustering

Recent advances in genomics and sequencing platforms have revolutionized our ability to create immense data sets, particularly for studying epigenetic regulation of gene expression. However, the avalanche of epigenomic data is difficult to parse for biological interpretation given nonlinear complex patterns and relationships. This attractive challenge in epigenomic data lends itself to machine learning for discerning infectivity and susceptibility. In this study, we explore over 3000 epigenomes of uninfected individuals and provide a framework to characterize the relationships among epigenetic modifiers, their modifiers, genetic loci, and specific immune cell types across all chromosomes using hierarchical clustering. Hierarchical clustering of epigenomic data revealed consistent epigenetic patterns across chromosomes, demonstrating that variation due to epigenetic modifiers is greater than variation between cell types. Gene Ontology and KEGG pathway analyses indicated significant enrichment of genes involved in chromatin remodeling, mRNA splicing, immune responses, and the regulation of microRNAs and snoRNAs. Epigenetic modifiers frequently formed biologically relevant clusters, including the cohesin complex, RNA Polymerase II transcription factors, and PRC2 complex members. These clustering behaviors remained consistent across all chromosomes, supported by entropy analysis and high Adjusted Rand Index scores, indicating robust cross-chromosomal similarity. Co-occurrence analysis further revealed specific sets of modifiers that consistently appeared together within clusters, reflecting shared biological functions and interactions. Validation using another dataset confirmed the reproducibility of these clustering patterns and modifier co-occurrence relationships, underscoring the reliability and generalizability of the methodology.

97 MATHEMATICS AND COMPUTING↗

Demographic Microsimulator for Integrated Urban Systems: Adapting Panel Survey of Income Dynamics to Capture the Continuum of Life

Agent-based models (ABMs) in transportation modeling simulate activity and travel decisions at the disaggregate level of households and individuals. To do this, ABMs require detailed and realistic information on agents’ socioeconomic and demographic characteristics. Various synthetic population generators have been proposed to address this need. However, most of those currently in practice are cross-sectional in nature and do not account for the dynamics within households and individuals as they progress through life events over time. This is a major shortcoming, as literature has shown that transportation decisions are affected by the transition between and co-occurrence of life cycle events. While some demographic evolution simulators have been proposed to address this issue, they are developed using cross-sectional data and capture only a small set of life cycle events and their interdependence. Addressing these drawbacks, we propose a demographic microsimulator (DEMOS) that captures the “continuum of life” by considering a range of household- and individual-level life cycle events. DEMOS is developed using the Panel Survey of Income Dynamics, one of the world’s longest-running longitudinal surveys. The DEMOS submodels consider key life cycle events that are influenced by agents’ demographic variables. DEMOS is applied to evolve the population of the San Francisco Bay Area over a 9-year horizon. Results demonstrate how DEMOS generates life trajectories and how DEMOS outputs match the observed demographic trends. DEMOS is expected to enable longitudinal analysis in the context of ABMs and expand ABMs analyses relating to dynamic processes such as household-level vehicle transactions.

Demographic evolution↗

Spectral and textural processing of ERTS imagery

A procedure is developed to simultaneously extract textural features from all bands of ERTS multispectral scanner imagery for automatic analysis. Multi-images lead to excessively large grey tone N-tuple co-occurrence matrices; therefore, neighboring grey N-tuple differences are measured and an ellipsoidally symmetric functional form is assumed for the co-occurrence distribution of multiimage greytone N-tuple differences. On the basis of past data the ellipsoidally symmetric approximation is shown to be reasonable. Initial evaluation of the procedure is encouraging.

Haralick, R. M.↗

Use of feature extraction techniques for the texture and context information in ERTS imagery: Spectral and textural processing of ERTS imagery

The author has identified the following significant results. A procedure was developed to extract cross-band textural features from ERTS MSS imagery. Evolving from a single image texture extraction procedure which uses spatial dependence matrices to measure relative co-occurrence of nearest neighbor grey tones, the cross-band texture procedure uses the distribution of neighboring grey tone N-tuple differences to measure the spatial interrelationships, or co-occurrences, of the grey tone N-tuples present in a texture pattern. In both procedures, texture is characterized in such a way as to be invariant under linear grey tone transformations. However, the cross-band procedure complements the single image procedure by extracting texture information and spectral information contained in ERTS multi-images. Classification experiments show that when used alone, without spectral processing, the cross-band texture procedure extracts more information than the single image texture analysis. Results show an improvement in average correct classification from 86.2% to 88.8% for ERTS image no. 1021-16333 with the cross-band texture procedure. However, when used together with spectral features, the single image texture plus spectral features perform better than the cross-band texture plus spectral features, with an average correct classification of 93.8% and 91.6%, respectively.

Haralick, R. H.↗

Predicting Drug Effects from High-dimensional Asymmetric Drug Data Sets using Graph Neural Networks: A Comprehensive Analysis of Multi-target Drug Effect Prediction

Graph neural networks (GNNs) have emerged as one of the most effective Machine learning (ML) techniques for drug effect prediction from drug molecular graphs. Despite having immense potential, GNN models lack performance when using data sets that contain high dimensional asymmetrically co-occurrent drug effects as targets with complex correlations between them. Training individual learning models for each drug effect and incorporating every prediction result for a wide spectrum of drug effects is beyond practicality. Such an implication provides a testbed to address this challenge as multi-target prediction problems, aiming to predict all drug effects at a time. We develop standard and hybrid graph neural networks (GNNs)to perform two separate tasks that are multi-regression for continuous values and multi-label classification for categorical values contained in our data sets. Since this step makes the target data even more sparse and introduces asymmetric label co-occurrence, the learning of multi-label classification models becomes difficult and heavily impacts the GNN's performance. To address these challenges, we propose a new data oversampling technique to improve multi-label classification performances on all the given imbalanced molecular graph data sets. Using the technique, we improve the data imbalance ratio of the drug effects better than before while protecting the data set's integrity. Finally, we evaluate multi-label classification performance using the best-performant hybrid GNN model on all the oversampled data sets obtained from the proposed oversampling technique. These results outperform those of other ML models including GNN models when they are trained on the original data sets or oversampled data sets using MLSMOTE (a well-known oversampling technique) in all evaluation metrics precision, recall, and F1 score by a significant margin.

Bose, Avishek [ORNL]↗

Breeding of microbiomes conferring salt tolerance to plants

Microbiome breeding through host-mediated selection is a technique to artificially select for microbiomes conferring beneficial properties to plants. Using a systematic selection protocol that maximises the heritability of microbiome effects, transmission fidelity, and microbiome stability through multiple selection cycles, we previously developed root-associated microbial communities conferring sodium and aluminium tolerance to Brachypodium distachyon, a model for cereal crops. Here, we explore the physiological mechanisms underlying our selected microbiomes’ effect on plant fitness and analyse how our selection protocol shaped the composition and structure of these microbiomes. We analysed the effects of our selected microbiomes on plant fitness and tissue-nutrient concentration, then used 16S rRNA amplicon sequencing to examine microbial community composition and co-occurrence network patterns. Our sodium-selected microbiomes reduced leaf sodium concentration by ~ 50%, whereas the aluminium-selected microbiomes had no effect on leaf-tissue nutrient concentration, suggesting different mechanisms underlying the microbiome-mediated stress tolerance. By testing the selected microbiomes in a cross-fostering experiment, we show that our artificially selected microbiomes attained (a) ecological robustness contributing to transplantability (i.e. inheritance) of microbiome-encoded effects between plants; and (b) network features identifying key bacteria promoting salt-stress tolerance. Combined, these findings elucidate critical mechanisms underlying host-mediated artificial selection as a framework to breed microbiomes with targeted benefits for plants under salt stresses, with significant implications for sustainable agriculture.

59 BASIC BIOLOGICAL SCIENCES↗

Insights into convergent evolution of cosexuality in liverworts from the Marchantia quadrata genome

Sex chromosomes are expected to coevolve with their respective sex, potentially disfavoring their co-occurrence as cosexuality evolves. This effect is expected to be stronger where sex chromosomes are restricted to one sex, such as in plants expressing sex in their haploid stage. We assess this hypothesis in liverworts with U/V sex chromosomes, ancestral dioicy, and several independent transitions to monoicy (cosexuality). We report the chromosome-level genome assembly of Marchantia quadrata, which recently evolved monoicy, and perform comparative genomic analyses with its dioicous relative M. polymorpha. We find that monoicy evolved via retention of the V chromosome as a small ninth chromosome, complete loss of the U chromosome, and translocation of key U-linked genes to autosomes, among which the major sex-determining gene (Feminizer) acquired environmental/developmental regulation. Our findings parallel recent observations on Ricciocarpos natans, which evolved monoicy independently, suggesting genetic constraints that may make transitions to monoicy predictable in liverworts.

Potente, Giacomo↗

A conserved chaperone protein is required for the formation of a noncanonical type VI secretion system spike tip complex

Type VI secretion systems (T6SSs) are dynamic protein nanomachines found in Gram-negative bacteria that deliver toxic effector proteins into target cells in a contact-dependent manner. Prior to secretion, many T6SS effector proteins require chaperones and/or accessory proteins for proper loading onto the structural components of the T6SS apparatus. However, despite their established importance, the precise molecular function of several T6SS accessory protein families remains unclear. In this study, we set out to characterize the DUF2169 family of T6SS accessory proteins. Using gene co-occurrence analyses, we find that DUF2169-encoding genes strictly co-occur with genes encoding T6SS spike complexes formed by valine-glycine repeat protein G (VgrG) and DUF4150 domains. Although structurally similar to Pro-Ala-Ala-Arg (PAAR) domains, “PAAR-like” DUF4150 domains lack PAAR motifs and instead contain a conserved PIPY motif, leading us to designate them PIPY domains. Next, we present both genetic and biochemical evidence that PIPY domains require a cognate DUF2169 protein to form a functional T6SS spike complex with VgrG. This contrasts with canonical PAAR proteins, which bind VgrG on their own to form functional spike complexes. By solving the first crystal structure of a DUF2169 protein, we show that this T6SS accessory protein adopts a novel protein fold. Furthermore, biophysical and structural modeling data suggest that DUF2169 contains a dynamic loop that physically interacts with a hydrophobic patch on the surface of its cognate PIPY domain. Based on these findings, we propose a model whereby DUF2169 proteins function as molecular chaperones that maintain VgrG–PIPY spike complexes in a secretion-competent state prior to their export by the T6SS apparatus.

DUF2169↗

DOME: Directional medical embedding vectors from Electronic Health Records

Motivation: The increasing availability of Electronic Health Record (EHR) systems has created enormous potential for translational research. Recent developments in representation learning techniques have led to effective large-scale representations of EHR concepts along with knowledge graphs that empower downstream EHR studies. However, most existing methods require training with patient-level data, limiting their abilities to expand the training with multi-institutional EHR data. On the other hand, scalable approaches that only require summary-level data do not incorporate temporal dependencies between concepts. Methods: We introduce a DirectiOnal Medical Embedding (DOME) algorithm to encode temporally directional relationships between medical concepts, using summary-level EHR data. Specifically, DOME first aggregates patient-level EHR data into an asymmetric co-occurrence matrix. Then it computes two Positive Pointwise Mutual Information (PPMI) matrices to correspondingly encode the pairwise prior and posterior dependencies between medical concepts. Following that, a joint matrix factorization is performed on the two PPMI matrices, which results in three vectors for each concept: a semantic embedding and two directional context embeddings. They collectively provide a comprehensive depiction of the temporal relationship between EHR concepts. Results: We highlight the advantages and translational potential of DOME through three sets of validation studies. First, DOME consistently improves existing direction-agnostic embedding vectors for disease risk prediction in several diseases, for example achieving a relative gain of 5.5% in the area under the receiver operating characteristic (AUROC) for lung cancer. Second, DOME excels in directional drug-disease relationship inference by successfully differentiating between drug side effects and indications, correspondingly achieving relative AUROC gain over the state-of-the-art methods by 10.8% and 6.6%. Finally, DOME effectively constructs directional knowledge graphs, which distinguish disease risk factors from comorbidities, thereby revealing disease progression trajectories. The source codes are provided at https://github.com/celehs/Directional-EHRembedding.

60 APPLIED LIFE SCIENCES↗

Structural uniformity and compositional homogeneity of solid-phase alloyed rod

Solid-phase processes have emerged as an alternative to fusion-based alloying to avoid coarse microstructures, undesirable phase formation, and high energy consumption. However, achieving uniform distribution of alloying elements during friction-based processing remains challenging due to highly heterogeneous thermomechanical conditions. This work evaluates the structural uniformity and compositional homogeneity of Al–Cu–Zn alloyed rods produced by friction extrusion (FE) and establishes the role of the rotational speed to feed rate ratio (N/V) on alloying effectiveness. A systematic matrix of FE experiments was conducted at constant extrusion ratio with N/V values ranging from 3.7 to 300. Compositional uniformity was assessed along the rod length (ICP-OES), in three dimensions (X-ray computed tomography), and at the microscale (SEM–EDS), supported by a gray-level co-occurrence matrix (GLCM)–based homogeneity metric. Smoothed particle hydrodynamics (SPH) simulations were used to reveal material flow and thermomechanical fields. Results show that N/V = 100 produces a high-shear mixing zone that eliminates the unmixed core and enables near-full dissolution and dispersion of Cu and Zn. At lower N/V, a laminar flow region persists at the rod center, causing segregation and large composition gradients. The combined experimental–computational analysis provides mechanistic insight into the transition from fragmented particle dispersion to thermomechanically assisted metallurgical mixing. This study establishes processing–structure relationships for solid-phase alloying and provides guidance for achieving homogenized compositions comparable to wrought alloys via rapid, scalable FE processing.

Aluminum↗

Complex Adsorption Behavior of Neodymium and Ytterbium on Structurally-Distinct Alumina Surfaces

New sources of rare earth elements (REEs) are needed to support a green energy transition. REEs adsorbed to aluminum-rich clays in weathering deposits represent important resources but the mechanisms responsible for their retention and ease of extraction are unresolved. Disordered coordination and co-occurrence of multiple species pose challenges to investigating REE adsorption processes via established spectroscopic methods. In this study, we applied element-specific surface crystallography methods to obtain a new perspective on the complexity of REE adsorption mechanisms and affinities. Alumina (001) and (012) crystal surfaces were utilized to evaluate surface-specific controls on Nd(III) and Yb(III) adsorption behavior. The REEs displayed similar total adsorption to alumina (001) as a mixture of inner- and outer-sphere complexes, but Nd displayed a greater proportion of inner-sphere binding. Adsorption of ordered inner- and outer-sphere REE species was substantially lower on alumina (012). These distinct behaviors reflect differences in the surface functional group charging and topography of the two surfaces. However, alumina (012) also hosted a substantial population of disordered adsorbed species, especially for Nd, potentially associated with Al vacancy surface defects. Here, the accumulation of light versus heavy REEs via adsorption in weathering deposits likely results from multiple, competing reactions affected by clay particle morphology. Leaching procedures for resource recovery should account for differential rates of desorption by coexisting inner- and outer-sphere REE surface complexes.

58 GEOSCIENCES↗

Colistin resistance plasmids dually enhance bacterial virulence and antibiotic resistance via surface polysaccharide biosynthesis

Plasmids carrying the mobilized colistin-resistance gene mcr-1 are prevalent among multidrug-resistant Gram-negative pathogens, yet their broad impact on bacterial physiology and virulence remains unclear. Here, we demonstrate that acquisition of an mcr-1 plasmid concurrently increases antimicrobial resistance and pathogenicity in Escherichia coli. On the same plasmid, the XRE-family transcriptional regulator EcaR cooperates with MCR-1 to activate the wec operon, driving biosynthesis of two surface polysaccharides: enterobacterial common antigen (ECA) and a high-molecular-weight O-chain. Expression of these surface polysaccharides increases bile resistance and virulence in a murine model and further elevates colistin resistance. MCR-1 enhances transcription of upstream genes in the wec operon, whereas EcaR directly activates an internal promoter (PwecE) to induce downstream gene expression. Thus, both components are required for surface polysaccharide expression, and deletion of either abolishes the phenotype. Genomic analysis of publicly available mcr plasmids reveals widespread co-occurrence of mcr-1 and ecaR on IncI2 and IncX4 plasmids, indicating their functional complementarity. These findings uncover a mechanism by which resistance plasmids remodel the bacterial surface, linking horizontal gene transfer to coordinated regulation of antimicrobial resistance and virulence.

Antimicrobial resistance↗

Vicennial metagenomic time series unveils evolutionary dynamics of giant viruses in a freshwater ecosystem

Giant viruses play crucial ecological roles in aquatic ecosystems, yet their evolutionary dynamics in response to environmental changes, particularly in freshwater environments, are not well understood. We analyzed a 20-year time series (2000-2019) of 471 co-assembled metagenomes from Lake Mendota (USA) to reconstruct 1512 giant virus metagenome-assembled genomes, providing insights into viral genome evolution. Viruses in the order Imitervirales dominate the virome, remaining consistent across seasons and years. Our findings reveal gene duplication (23% of genes) and horizontal gene transfer (29% of genes) as key drivers of genomic innovation. A co-occurrence network analysis indicates increased virus-host interactions following the introduction of an invasive predatory zooplankton in 2009, highlighting potential hosts in Bigyra, Perkinsea, and Euglenozoa. While single nucleotide polymorphism analysis shows predominantly purifying selection in viral genes, there is a significant increase in positively selected genes post-invasion, particularly those related to infection. Comparative evolutionary analyses reveal that giant viruses exhibit genome-wide substitution rates similar to co-occurring bacteria but significantly slower than smaller dsDNA phages, suggesting both stability and adaptability. Our study demonstrates that freshwater giant viruses employ various evolutionary strategies to respond to environmental change. These results underscore their significant yet often underappreciated role in freshwater ecosystem dynamics.

Vasquez, Yumary M↗

Seasonal Host Shifts for Legionella Within an Industrial Water‐Cooling System

Legionella is a genus of environmental bacteria containing pathogenic species such as Legionella pneumophila that are responsible for Legionnaires' disease, a potentially fatal respiratory infection. Disease aetiology can involve Legionella replication intracellularly within protists and this study aimed to characterise the Legionella -protist relationship to develop novel outbreak prevention targets. Water and sediment samples were collected from a water-cooling tower in South Carolina over a 6-month period. Concomitantly, multiple environmental parameters were recorded. Bacterial and eukaryotic communities were characterised using 16S rRNA gene V4 region and a 252 bp fragment of 18S rRNA gene, respectively. Co-occurrence network analyses were performed to elucidate Legionella -protist correlations through time. We found that Legionella correlated with different protists as the seasons progressed. Acanthamoeba correlated with Legionella in early spring followed by Vannella and Korotnevella in late spring and early summer, and were joined by Echinamoeba in mid-summer. Vannella and Acanthamoeba are known potential hosts for Legionella , while Korotnevella is a potential undocumented host. Of the environmental parameters, temperature showed strong correlation with protists genera, suggesting that Legionella abundance was driven by temperature-dependent protist availability. Our results highlight ecological shifts that are associated with elevated Legionella levels, which offers potential targets to help predict and prevent disease outbreaks.

microbial communities↗