Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Subgroup analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

49 records · Page 3

Multiscale X-ray phase contrast imaging of human cartilage for investigating osteoarthritis formation

The evolution of cartilage degeneration is still not fully understood, partly due to its thinness, low radio-opacity and therefore lack of adequately resolving imaging techniques. X-ray phase-contrast imaging (X-PCI) offers increased sensitivity with respect to standard radiography and CT allowing an enhanced visibility of adjoining, low density structures with an almost histological image resolution. This study examined the feasibility of X-PCI for high-resolution (sub-) micrometer analysis of different stages in tissue degeneration of human cartilage samples and compare it to histology and transmission electron microscopy. Ten 10%-formalin preserved healthy and moderately degenerated osteochondral samples, post-mortem extracted from human knee joints, were examined using four different X-PCI tomographic set-ups using synchrotron radiation the European Synchrotron Radiation Facility (France) and the Swiss Light Source (Switzerland). Volumetric datasets were acquired with voxel sizes between 0.7 × 0.7 × 0.7 and 0.1 × 0.1 × 0.1 µm 3 . Data were reconstructed by a filtered back-projection algorithm, post-processed by ImageJ, the WEKA machine learning pixel classification tool and VGStudio max. For correlation, osteochondral samples were processed for histology and transmission electron microscopy. X-PCI provides a three-dimensional visualization of healthy and moderately degenerated cartilage samples down to a (sub-)cellular level with good correlation to histologic and transmission electron microscopy images. X-PCI is able to resolve the three layers and the architectural organization of cartilage including changes in chondrocyte cell morphology, chondrocyte subgroup distribution and (re-)organization as well as its subtle matrix structures. X-PCI captures comprehensive cartilage tissue transformation in its environment and might serve as a tissue-preserving, staining-free and volumetric virtual histology tool for examining and chronicling cartilage behavior in basic research/laboratory experiments of cartilage disease evolution.

3D analysis↗

Recommendations for Distributed Energy Resource Access Control

Cybersecurity for internet - connected Distributed Energy Resources (DER) is essential for the safe and reliable operation of the US power system. Many facets of DER cybersecurity are currently being investigated within different standards development organizations, research communities, and industry committees to address this critical need. This report covers DER access control guidance compiled by the Access Controls Subgroup of the SunSpec/Sandia DER Cybersecurity Workgroup. The goal of the group was to create a consensus - based technical framework to minimize the risk of unauthorized access to DER systems. The subgroup set out to define a strict control environment where users are authorized to access DER monitoring and control features through three steps: (a) user is identified using a proof-of-identity, (b) the user is authenticated by a managed database, (c) and the user is authorized for a specific level of access. DER access control also provides accountability and nonrepudiation within the power system control environment that can be used for forensic analysis and attribution in the event of a cyber-attack. This paper covers foundational requirements for a DER access control environment as well as offering a collection of possible policy, model, and mechanism implementation approaches for IEEE 1547-mandated communication protocols.

45 MILITARY TECHNOLOGY, WEAPONRY, AND NATIONAL DEF↗

Heavy-tailed distribution of the number of papers within scientific journals

Scholarly publications represent at least two benefits for the study of the scientific community as a social group. First, they attest to some form of relation between scientists (collaborations, mentoring, heritage, …), useful to determine and analyze social subgroups. Second, most of them are recorded in large databases, easily accessible and including a lot of pertinent information, easing the quantitative and qualitative study of the scientific community. Understanding the underlying dynamics driving the creation of knowledge in general, and of scientific publication in particular, can contribute to maintaining a high level of research, by identifying good and bad practices in science. In this article, we aim to advance this understanding by a statistical analysis of publication within peer-reviewed journals. Namely, we show that the distribution of the number of papers published by an author in a given journal is heavy-tailed, but has a lighter tail than a power law. Interestingly, we demonstrate (both analytically and numerically) that such distributions match the result of a modified preferential attachment process, where, on top of a Barabási-Albert process, we take the finite career span of scientists into account.

96 KNOWLEDGE MANAGEMENT AND PRESERVATION↗

Initial Efforts Organizing WPNCS SG-8: Preservation of Expert Knowledge and Judgement Applied to Criticality Benchmarks

The Working Party on Nuclear Criticality Safety (WPNCS) under the guidance of the Organization for Economic Co-operation and Development (OECD) Nuclear Energy Agency (NEA) has over 20 years of experience addressing concerns related to static and transient configurations encountered within the nuclear fuel cycle: fuel fabrication, transportation, reprocessing, storage, and geological disposal. One of the cornerstone activities of the WPNCS is the International Criticality Safety Benchmark Evaluation Project (ICSBEP), which was established to identify a comprehensive set of criticality benchmark data, evaluate the data, including quantification of overall uncertainties; compile the data into a standardized format, perform sample calculations utilizing modern nuclear data sets and codes utilized in nuclear criticality safety, and formally document the work into a single source of verified benchmark data. Annually, members of the ICSBEP Technical Review Group (TRG) contribute evaluated benchmark data that undergoes comprehensive technical review prior to publication in the ICSBEP Handbook. In the years since the ICSBEP was established, there has been much work to prepare benchmark data to support validation activities in nuclear criticality safety. The 2020 edition of the ICSBEP Handbook contains acceptable benchmark specifications for 5,053 critical, subcritical, or near-critical configurations in 582 benchmark evaluations. Modern benchmark development benefits from decades of experienced international participants, a well-established handbook format, supplementary guides to deal with uncertainty quantification, and a comprehensive review process based upon independent reviews from international experts. The ICSBEP Handbook also contains 838 configurations deemed unacceptable to support criticality safety efforts. They are recorded, with the reasoning for their rejection, to preserve the experimental data, prevent reevaluation of data that are incomplete or contain known errors, and/or to potentially allow future reevaluation of the experiment pending the identification of sufficient data to resolve identified inconsistencies and errors. Users of the ICSBEP Handbook today might notice that the rigor and quality of modern criticality safety benchmarks is much greater than those prepared within the initial decade of the project. Benchmarks with 1s uncertainties in k eff greater than 1% were traditionally rejected unless they were identified as unique experiment types that encompassed materials, fuels, or designs not available from other benchmark experiments. However, benchmarks developed using modern experimental techniques and practices typically have uncertainties on the order of a few tenths of a percent. There have been ongoing efforts to improve the overall quality of previously published benchmark evaluations. Seventy-eight evaluations, containing approximately 600 configurations, have been revised just within the past decade. An additional eleven benchmarks are under revision for updated release in the 2020 edition of the ICSBEP Handbook. If some of the historic benchmarks were resubmitted in their current form to the TRG today, they would be rejected due to lack of data, missing components in the uncertainty analysis, or incomplete benchmark model development. The use of historic criticality safety benchmarks that underestimate the total uncertainty, lack properly quantified biases, or provide inadequate benchmark specifications do not sufficiently support modern criticality safety and nuclear data efforts. Although the ICSBEP Handbook is recognized by regulating bodies to support criticality safety, users are required to justify their reasons to ignore historic benchmark data and include additional safety margins within their designs. Discussions were held at the WPNCS 23rd Annual Meeting in September 2019 regarding the aforementioned issues. The resultant decision was to establish Subgroup 8 (SG-8): Preservation of Expert Knowledge and Judgement Applied to Criticality Benchmarks. The current activities of SG-8 are discussed herein.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Measuring Equality in Machine Learning Security Defenses: A Case Study in Speech Recognition

Over the past decade, the machine learning security community has developed a myriad of defenses for evasion attacks. An understudied question in that community is: for whom do these defenses defend? This work considers common approaches to defending learned systems and how security defenses result in performance inequities across different sub-populations. We outline appropriate parity metrics for analysis and begin to answer this question through empirical results of the fairness implications of machine learning security methods. We find that many methods that have been proposed can cause direct harm, like false rejection and unequal benefits from robustness training. The framework we propose for measuring defense equality can be applied to robustly trained models, preprocessing-based defenses, and rejection methods. We identify a set of datasets with a user-centered application and a reasonable computational cost suitable for case studies in measuring the equality of defenses. In our case study of speech command recognition, we show how such adversarial training and augmentation have non-equal but complex protections for social subgroups across gender, accent, and age in relation to user coverage. We present a comparison of equality between two rejection-based defenses: randomized smoothing and neural rejection, finding randomized smoothing more equitable due to the sampling mechanism for minority groups. This represents the first work examining the disparity in the adversarial robustness in the speech domain and the fairness evaluation of rejection-based defenses.

• Artificial intelligence (AI) / machine learning ↗

Streptococcus pneumoniae HtrA is a dynamic and monomeric virulence factor capable of forming larger oligomeric complexes

Abstract High‐temperature requirement A (HtrA) proteases are a conserved family of serine proteases central to protein quality control and bacterial virulence. While Gram‐negative and human HtrAs are structurally well studied, Gram‐positive homologs remain essentially uncharacterized. Here, we present the first integrated structural and mechanistic analysis of a Gram‐positive HtrA, from Streptococcus pneumoniae , a virulence factor essential for adhesion and infection in vivo. Proteomic profiling of an htrA knockout and cleavage assays demonstrate that S. pneumoniae HtrA is required for protein quality control, with the PDZ domain mediating substrate recognition. Biochemically, S. pneumoniae HtrA exists exclusively as a monomer in solution, a striking divergence from canonical trimeric HtrAs that we show is shared with other Gram‐positive homologs. NMR analyses reveal that the monomer dynamically samples open and closed conformations, while cryo‐EM of a catalytic mutant identifies a hexamer stabilized by a unique LoopA–PDZ interaction. Together, these findings define S. pneumoniae HtrA as a dynamic monomer with interdomain coupling between its protease and PDZ domains, establishing Gram‐positive HtrAs as a mechanistically divergent subgroup within the HtrA family.

Lee, Eunjeong [Department of Biochemistry and Mole↗

Domain Shift Analysis in Chest Radiographs Classification in a Veterans Healthcare Administration Population

This study aims to assess the impact of domain shift on chest X-ray classification accuracy and to analyze the influence of ground truth label quality and demographic factors such as age group, sex, and study year. We used a DenseNet121 model pre-trained MIMIC-CXR dataset for deep learning-based multi-label classification using ground truth labels from radiology reports extracted using the CheXpert and CheXbert Labeler. We compared the performance of the 14 chest X-ray labels on the MIMIC-CXR and Veterans Healthcare Administration chest X-ray dataset (VA-CXR). The validation of ground truth and the assessment of multi-label classification performance across various NLP extraction tools revealed that the VA-CXR dataset exhibited lower disagreement rates than the MIMIC-CXR datasets. Additionally, there were notable differences in AUC scores between models utilizing CheXpert and CheXbert. When evaluating multi-label classification performance across different datasets, minimal domain shift was observed in the unseen VA dataset, except for the label “Enlarged Cardiomediastinum.” The subgroup with the most significant variations in multi-label classification performance was study year. These findings underscore the importance of considering domain shift in chest X-ray classification tasks, paying particular attention to the temporality of the exam. Our study reveals the significant impact of domain shift and demographic factors on chest X-ray classification, emphasizing the need for improved transfer learning and robust model development. Addressing these challenges is crucial for advancing medical imaging research and improving patient care.

chest X-ray image classification↗

GAL08, an Uncultivated Group of Acidobacteria, Is a Dominant Bacterial Clade in a Neutral Hot Spring

GAL08 are bacteria belonging to an uncultivated phylogenetic cluster within the phylum Acidobacteria . We detected a natural population of the GAL08 clade in sediment from a pH-neutral hot spring located in British Columbia, Canada. To shed light on the abundance and genomic potential of this clade, we collected and analyzed hot spring sediment samples over a temperature range of 24.2–79.8°C. Illumina sequencing of 16S rRNA gene amplicons and qPCR using a primer set developed specifically to detect the GAL08 16S rRNA gene revealed that absolute and relative abundances of GAL08 peaked at 65°C along three temperature gradients. Analysis of sediment collected over multiple years and locations revealed that the GAL08 group was consistently a dominant clade, comprising up to 29.2% of the microbial community based on relative read abundance and up to 4.7 × 10 5 16S rRNA gene copy numbers per gram of sediment based on qPCR. Using a medium quality threshold, 25 single amplified genomes (SAGs) representing these bacteria were generated from samples taken at 65 and 77°C, and seven metagenome-assembled genomes (MAGs) were reconstructed from samples collected at 45–77°C. Based on average nucleotide identity (ANI), these SAGs and MAGs represented three separate species, with an estimated average genome size of 3.17 Mb and GC content of 62.8%. Phylogenetic trees constructed from 16S rRNA gene sequences and a set of 56 concatenated phylogenetic marker genes both placed the three GAL08 bacteria as a distinct subgroup of the phylum Acidobacteria , representing a candidate order ( Ca. Frugalibacteriales) within the class Blastocatellia. Metabolic reconstructions from genome data predicted a heterotrophic metabolism, with potential capability for aerobic respiration, as well as incomplete denitrification and fermentation. In laboratory cultivation efforts, GAL08 counts based on qPCR declined rapidly under atmospheric levels of oxygen but increased slightly at 1% (v/v) O 2 , suggesting a microaerophilic lifestyle.

59 BASIC BIOLOGICAL SCIENCES↗

SPARC: Structural properties associated with residue constraints

SPARC facilitates the generation of plausible hypotheses regarding underlying biochemical mechanisms by structurally characterizing protein sequence constraints. Such constraints appear as residues co-conserved in functionally related subgroups, as subtle pairwise correlations (i.e., direct couplings), and as correlations among these sequence features or with structural features. SPARC performs three types of analyses. First, based on pairwise sequence correlations, it estimates the biological relevance of alternative conformations and of homomeric contacts, as illustrated here for death domains. Second, it estimates the statistical significance of the correspondence between directly coupled residue pairs and interactions at heterodimeric interfaces. Third, given molecular dynamics simulated structures, it characterizes interactions among constrained residues or between such residues and ligands that: (a) are stably maintained during the simulation; (b) undergo correlated formation and/or disruption of interactions with other constrained residues; or (c) switch between alternative interactions. We illustrate this for two homohexameric complexes: the bacterial enhancer binding protein (bEBP) NtrC1, which activates transcription by remodeling RNA polymerase (RNAP) containing σ 54 , and for DnaB helicase, which opens DNA at the bacterial replication fork. Based on the NtrC1 analysis, we hypothesize possible mechanisms for inhibiting ATP hydrolysis until ADP is released from an adjacent subunit and for coupling ATP hydrolysis to restructuring of σ 54 binding loops. Based on the DnaB analysis, we hypothesize that DnaB ‘grabs’ ssDNA by flipping every fourth base and inserting it into cavities between subunits and that flipping of a DnaB-specific glutamine residue triggers ATP hydrolysis.

97 MATHEMATICS AND COMPUTING↗

SN 2021fxy: mid-ultraviolet flux suppression is a common feature of Type Ia supernovae

ABSTRACT We present ultraviolet (UV) to near-infrared (NIR) observations and analysis of the nearby Type Ia supernova SN 2021fxy. Our observations include UV photometry from Swift/UVOT, UV spectroscopy from HST/STIS, and high-cadence optical photometry with the Swope 1-m telescope capturing intranight rises during the early light curve. Early B − V colours show SN 2021fxy is the first ‘shallow-silicon’ (SS) SN Ia to follow a red-to-blue evolution, compared to other SS objects which show blue colours from the earliest observations. Comparisons to other spectroscopically normal SNe Ia with HST UV spectra reveal SN 2021fxy is one of several SNe Ia with flux suppression in the mid-UV. These SNe also show blueshifted mid-UV spectral features and strong high-velocity Ca ii features. One possible origin of this mid-UV suppression is the increased effective opacity in the UV due to increased line blanketing from high velocity material, but differences in the explosion mechanism cannot be ruled out. Among SNe Ia with mid-UV suppression, SNe 2021fxy and 2017erp show substantial similarities in their optical properties despite belonging to different Branch subgroups, and UV flux differences of the same order as those found between SNe 2011fe and 2011by. Differential comparisons to multiple sets of synthetic SN Ia UV spectra reveal this UV flux difference likely originates from a luminosity difference between SNe 2021fxy and 2017erp, and not differing progenitor metallicities as suggested for SNe 2011by and 2011fe. These comparisons illustrate the complicated nature of UV spectral formation, and the need for more UV spectra to determine the physical source of SNe Ia UV diversity.

Astronomy & Astrophysics↗

ARCH: Large-scale knowledge graph via aggregated narrative codified health records analysis

Objective: Electronic health record (EHR) systems contain a wealth of clinical data stored as both codified data and free-text narrative notes (NLP). The complexity of EHR presents challenges in feature representation, information extraction, and uncertainty quantification. Here, to address these challenges, we proposed an efficient Aggregated naRrative Codified Health (ARCH) records analysis to generate a large-scale knowledge graph (KG) for a comprehensive set of EHR codified and narrative features. Methods: Using data from 12.5 million Veterans Affairs patients, ARCH first derives embedding vectors and generates similarities along with associated p-values to measure the strength of relatedness between clinical features with statistical certainty quantification. Next, ARCH performs a sparse embedding regression to remove indirect linkage between features to build a sparse KG. Finally, ARCH was validated on various clinical tasks, including detecting known relationships between entity pairs, predicting drug side effects, disease phenotyping, as well as sub-typing Alzheimer’s disease patients. Results: ARCH produces high-quality clinical embeddings and KG for over 60,000 codified and narrative EHR concepts. The KG and embeddings are visualized in the R-shiny powered web-API.3 ARCH achieved high accuracy in detecting EHR concept relationships, with AUCs of 0.926 (codified) and 0.861 (NLP) for similar EHR concepts, and 0.810 (codified) and 0.843 (NLP) for related pairs. It detected drug side effects with a 0.723 AUC, which improved to 0.826 after fine-tuning. Using both codified and NLP features, the detection power increased significantly. Compared to other methods, ARCH has superior accuracy and enhances weakly supervised phenotyping algorithms’ performance. Notably, it successfully categorized Alzheimer’s patients into two subgroups with varying mortality rates. Conclusion: The proposed ARCH algorithm generates large-scale high-quality semantic representations and knowledge graph for both codified and NLP EHR features, useful for a wide range of predictive modeling tasks.

Electronic health records↗

Substructure in the stellar halo near the Sun: I. Data-driven clustering in integrals-of-motion space

Context. Merger debris is expected to populate the stellar haloes of galaxies. In the case of the Milky Way, this debris should be apparent as clumps in a space defined by the orbital integrals of motion of the stars. Aims. Our aim is to develop a data-driven and statistics-based method for finding these clumps in integrals-of-motion space for nearby halo stars and to evaluate their significance robustly. Methods. We used data from Gaia EDR3, extended with radial velocities from ground-based spectroscopic surveys, to construct a sample of halo stars within 2.5 kpc from the Sun. We applied a hierarchical clustering method that makes exhaustive use of the single linkage algorithm in three-dimensional space defined by the commonly used integrals of motion energy E, together with two components of the angular momentum, L z and L ⊥ . To evaluate the statistical significance of the clusters, we compared the density within an ellipsoidal region centred on the cluster to that of random sets with similar global dynamical properties. By selecting the signal at the location of their maximum statistical significance in the hierarchical tree, we extracted a set of significant unique clusters. By describing these clusters with ellipsoids, we estimated the proximity of a star to the cluster centre using the Mahalanobis distance. Additionally, we applied the HDBSCAN clustering algorithm in velocity space to each cluster to extract subgroups representing debris with different orbital phases. Results. Our procedure identifies 67 highly significant clusters (> 3σ), containing 12% of the sources in our halo set, and 232 subgroups or individual streams in velocity space. In total, 13.8% of the stars in our data set can be confidently associated with a significant cluster based on their Mahalanobis distance. Inspection of the hierarchical tree describing our data set reveals a complex web of relations between the significant clusters, suggesting that they can be tentatively grouped into at least six main large structures, many of which can be associated with previously identified halo substructures, and a number of independent substructures. This preliminary conclusion is further explored in a companion paper, in which we also characterise the substructures in terms of their stellar populations. Conclusions. Our method allows us to systematically detect kinematic substructures in the Galactic stellar halo with a data-driven and interpretable algorithm. The list of the clusters and the associated star catalogue are provided in two tables available at the CDS.

79 ASTRONOMY AND ASTROPHYSICS↗

Intrinsic and environmental drivers of pairwise cohesion in wild Canis social groups

Animals within social groups respond to costs and benefits of sociality by adjusting the proportion of time they spend in close proximity to other individuals in the group (cohesion). Variation in cohesion between individuals, in turn, shapes important group-level processes such as subgroup formation and fission–fusion dynamics. Although critical to animal sociality, a comprehensive understanding of the factors influencing cohesion remains a gap in our knowledge of cooperative behavior in animals. We tracked 574 individuals from six species within the genus Canis in 15 countries on four continents with GPS telemetry to estimate the time that pairs of individuals within social groups spent in close proximity and test hypotheses regarding drivers of cohesion. Pairs of social canids (Canis spp.) varied widely in the proportion of time they spent together (5%–100%) during seasonal monitoring periods relative to both intrinsic characteristics and environmental conditions. The majority of our data came from three species of wolves (gray wolves, eastern wolves, and red wolves) and coyotes. For these species, cohesion within social groups was greatest between breeding pairs and varied seasonally as the nature of cooperative activities changed relative to annual life history patterns. Across species, wolves were more cohesive than coyotes. For wolves, pairs were less cohesive in larger groups, and when suitable, small prey was present reflecting the constraints of food resources and intragroup competition on social associations. Pair cohesion in wolves declined with increased anthropogenic modification of the landscape and greater climatic variability, underscoring challenges for conserving social top predators in a changing world. We show that pairwise cohesion in social groups varies strongly both within and across Canis species, as individuals respond to changing ecological context defined by resources, competition, and anthropogenic disturbance. Our work highlights that cohesion is a highly plastic component of animal sociality that holds significant promise for elucidating ecological and evolutionary mechanisms underlying cooperative behavior.

59 BASIC BIOLOGICAL SCIENCES↗