Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “robust clustering”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 127 records · Page 7

Ligand-Driven Electrochemical Tuning of Co 6 Se 8 Chevrel Clusters

Molecular “Chevrel-type” clusters of the formula Co 6 Se 8 L 6 (L = neutral ligand) are a well-studied class of clusters due to their utility as molecular analogues to the Chevrel extended solid phase and their application as subunits in hierarchical materials. However, their solution and optical properties remain relatively underexplored. Aiming to develop the fundamental relationships between the molecular and electronic structures of these clusters and their electrochemical and photophysical properties, this work reports the preparation of a series of Co 6 Se 8 (P­(C 6 H 4 R) 3 ) 6 -type clusters with R = Cl (1), F (2), H (3), CH 3 (4), and OCH 3 (5) via a stepwise synthetic approach. Solution and solid-state experimental characterization and density functional theory calculations reveal that the Co 6 Se 8 cores of 1–5 maintain consistent electronic and structural properties despite the variation of the triarylphosphine ligand para-substituent Hammett parameters (σ p ). However, cyclic voltammetry measurements indicate that the electron transfer energetics of 1–5 are strongly influenced by ligand substitution, with the E 1/2 of a given redox event spanning ∼0.5 V depending on the triarylphosphine ligand’s σ p . In conclusion, these findings support the characterization of Co 6 Se 8 clusters as atomically precise nanoclusters with both the structural robustness and the electrochemical tunability needed to act as components in larger charge transfer assemblies.

Wheaton, Amelia M. [Argonne National Laboratory (A↗

Optical properties of a diamond NV color center from capped embedded multiconfigurational correlated wavefunction theory

Diamond defects are among the most promising qubits. Modeling their properties through accurate quantum mechanical simulations can further their development into robust units of information. We use the recently developed capped density functional embedding theory (capped-DFET) with the multiconfigurational n-electron valence second-order perturbation theory to characterize the electronic excitation energies for different spin manifolds of the well-characterized negatively charged substitutional N defect adjacent to a vacancy (V C ) in diamond (N C V C − ). We successfully reproduce vertical excitation energies for both triplet and singlet states of N C V C − with errors < 0.1 eV. Unlike other embedding methods, capped-DFET exhibits robust predictions that are approximately independent of the embedded cluster size: it only requires a cluster to contain the defect atoms and their nearest neighbors (as small as a 40-atom capped cluster). Furthermore, our method is free from slowly converging Coulomb interactions between charged defects, and thus also only weakly dependent on supercell size.

Chemistry↗

Leveraging BERT and Network-Based Attention Analysis for Identifying Treatment Milestones in EHRs

This study introduces a sophisticated data-driven framework for analyzing Electronic Health Records (EHRs) using transformer-based models to identify and disentangle overlapping treatment contexts. The framework leverages a preprocessing pipeline that transforms structured procedural codes into semantically enriched descriptive text, enabling the use of attention mechanisms to cluster medical events into treatment milestones—cohesive and distinct components of care processes. The methodology is rigorously validated using synthetic datasets derived from the MIMIC-III database, designed to simulate the heterogeneity and overlapping procedural contexts characteristic of real-world EHR scenarios. Quantitative evaluation highlights the framework’s robustness in disentangling concurrent care pathways, with attention metrics and unsupervised clustering approaches demonstrating the ability to preserve intra-context relationships while distinguishing inter-context dependencies. By addressing challenges inherent in data heterogeneity, this approach provides a foundation for uncovering complex treatment patterns, advancing clinical decision-making, and optimizing resource allocation in diverse healthcare environments.

Kim, Minsu [ORNL] (ORCID:0000000224185535)↗

Neutrino event selection in the MicroBooNE liquid argon time projection chamber using Wire-Cell 3D imaging, clustering, and charge-light matching

An accurate and efficient event reconstruction is required to realize the full scientific capability of liquid argon time projection chambers (LArTPCs). The current and future neutrino experiments that rely on massive LArTPCs create a need for new ideas and reconstruction approaches. Wire-Cell, proposed in recent years, is a novel tomographic event reconstruction method for LArTPCs. The Wire-Cell 3D imaging approach capitalizes on charge, sparsity, time, and geometry information to reconstruct a topology-agnostic 3D image of the ionization electrons prior to pattern recognition. A second novel method, the many-to-many charge-light matching, then pairs the TPC charge activity to the detected scintillation light signal, thus enabling a powerful rejection of cosmic-ray muons in the MicroBooNE detector. A robust processing of the scintillation light signal and an appropriate clustering of the reconstructed 3D image are fundamental to this technique. In this paper, we describe the principles and algorithms of these techniques and their successful application in the MicroBooNE experiment. A quantitative evaluation of the performance of these techniques is presented. Using these techniques, a 95% efficient pre-selection of neutrino charged-current events is achieved with a 30-fold reduction of non-beam-coincident cosmic-ray muons, and about 80% of the selected neutrino charged-current events are reconstructed with at least 70% completeness and 80% purity.

3D imaging↗

Three dimensional cluster analysis for atom probe tomography using Ripley’s K-function and machine learning

The size and structure of spatial molecular and atomic clustering can significantly impact material properties and is therefore important to accurately quantify. Ripley’s K-function (K(r)), a measure of spatial correlation, can be used to perform such quantification when the material system of interest can be represented as a marked point pattern. This work demonstrates how machine learning models based on K (r)-derived metrics can accurately estimate cluster size and intra-cluster density in simulated three dimensional (3D) point patterns containing spherical clusters of varying size; over 90% of model estimates for cluster size and intra-cluster density fall within 11% and 18% error of the true values, respectively. These K (r)-based size and density estimates are then applied to an experimental APT reconstruction to characterize MgZn clusters in a 7000 series aluminum alloy. Here we find that the estimates are more accurate, consistent, and robust to user interaction than estimates from the popular maximum separation algorithm. Using K (r) and machine learning to measure clustering is an accurate and repeatable way to quantify this important material attribute.

36 MATERIALS SCIENCE↗

Statistical relationships across epigenomes using large-scale hierarchical clustering

Recent advances in genomics and sequencing platforms have revolutionized our ability to create immense data sets, particularly for studying epigenetic regulation of gene expression. However, the avalanche of epigenomic data is difficult to parse for biological interpretation given nonlinear complex patterns and relationships. This attractive challenge in epigenomic data lends itself to machine learning for discerning infectivity and susceptibility. In this study, we explore over 3000 epigenomes of uninfected individuals and provide a framework to characterize the relationships among epigenetic modifiers, their modifiers, genetic loci, and specific immune cell types across all chromosomes using hierarchical clustering. Hierarchical clustering of epigenomic data revealed consistent epigenetic patterns across chromosomes, demonstrating that variation due to epigenetic modifiers is greater than variation between cell types. Gene Ontology and KEGG pathway analyses indicated significant enrichment of genes involved in chromatin remodeling, mRNA splicing, immune responses, and the regulation of microRNAs and snoRNAs. Epigenetic modifiers frequently formed biologically relevant clusters, including the cohesin complex, RNA Polymerase II transcription factors, and PRC2 complex members. These clustering behaviors remained consistent across all chromosomes, supported by entropy analysis and high Adjusted Rand Index scores, indicating robust cross-chromosomal similarity. Co-occurrence analysis further revealed specific sets of modifiers that consistently appeared together within clusters, reflecting shared biological functions and interactions. Validation using another dataset confirmed the reproducibility of these clustering patterns and modifier co-occurrence relationships, underscoring the reliability and generalizability of the methodology.

97 MATHEMATICS AND COMPUTING↗

Robust Control Design for Systems With Probabilistic Uncertainty

This paper presents a reliability- and robustness-based formulation for robust control synthesis for systems with probabilistic uncertainty. In a reliability-based formulation, the probability of violating design requirements prescribed by inequality constraints is minimized. In a robustness-based formulation, a metric which measures the tendency of a random variable/process to cluster close to a target scalar/function is minimized. A multi-objective optimization procedure, which combines stability and performance requirements in time and frequency domains, is used to search for robustly optimal compensators. Some of the fundamental differences between the proposed strategy and conventional robust control methods are: (i) unnecessary conservatism is eliminated since there is not need for convex supports, (ii) the most likely plants are favored during synthesis allowing for probabilistic robust optimality, (iii) the tradeoff between robust stability and robust performance can be explored numerically, (iv) the uncertainty set is closely related to parameters with clear physical meaning, and (v) compensators with improved robust characteristics for a given control structure can be synthesized.

Crespo, Luis G.↗

Uneven Inflation Load Share Trends in Clusters

The use of parachute clusters for payload recovery is still seeing widespread use ever since the early days of WWII. By involving the (near) simultaneous deployment of several smaller and identical canopies connected to the payload, cluster systems offer flexibility in tailoring to needed descent rates and load management, as well as providing robustness against individual canopy deployment or opening failure. Their downside, of course, resides in the possibility of differing inflation rates by each cluster member as caused by deployment variability, canopy-to-canopy interference, etc. Such variability leads to the lead-lag phenomenon, which causes uneven loading among the parachutes, often times leaving a single canopy to take up a significant portion of the system’s inflation loads. Herein we investigate how serious such an effect can be in terms of the number N of cluster members, underinflation drag of the lagging canopies, inflation swiftness of the leader in comparison to the laggards’, disreefing cutter activation staggering and pre-disreefing drag area. Two new metrics are used to highlight load share unevenness, namely, the peak and average leader canopy riser load in comparison to the leader’s drag during no-lead-lag; and leader peak load, as compared to total peak load. Results are calculated from data collected in NASA’s Orion/CPAS test program, as well as from simple algebraic expressions informing leader drag as sustained in different cluster systems (i.e., of different N) and varying leader inflation time relative to the laggards. Generally, using large-N cluster systems confers better load sharing among canopies. However, and in deployments where significant lead-lag occur, large-N systems may feature greater leader overload excursions, oftentimes in excess of 50% the no-lead-lag levels. These excursions are particularly made worse when cutter activation among the members are far from simultaneous and the laggards’ pre-disreefing drag area is small in comparison to the leader’s.

parachutes↗

Stochastic economic dispatch of wind power under uncertainty using clustering-based extreme scenarios

Operation of power systems with high penetrations of renewable energy sources requires tools for robust decision making under uncertainty. Stochastic economic dispatch and stochastic unit commitment are effective techniques for planning and operation under uncertainty, whose effectiveness depends on the cardinality and quality of the scenario set. Here, this article proposes a machine learning method using -means clustering for capturing relevant physical information from a large population of analog scenarios. Extreme scenario samples drawn from the clusters are used in a two-stage stochastic economic dispatch computation. The effectiveness of the proposed approach is assessed on a synthetic 200-bus system with a geographic footprint over Illinois, USA for four months from each season of WIND Toolkit data. The combination of -means clustering with importance sampling is shown to reduce the total operational cost by over 43% compared to sampling from populations based on heuristic clustering-based methods. Additionally, the variability in the mean cost is about 56% lower than the variability using Monte Carlo sampling. Moreover, the operational cost with the presented approach is shown to be close to the cost calculated based on a hindsight exact wind profile, signifying a highly accurate quantification of wind uncertainty by the presented -means clustering based sampling method.

17 WIND ENERGY↗

Convergence rate enhancement of navier-stokes codes on clustered grids

Our Sensitivity-Based Minimal Residual (SBMR) method which is based on our earlier Distributed Minimal Residual (DMR) method allows each component of the solution vector in a system of equations to have its own convergence speed. Our global SBMR method was found to consistently outperform the DMR method while requiring considerably less computer memory. Recently, we have developed and tested a new Line SBMR or LSBMR method and a Time-Step-Scaling (TSS) method that are even more robust and computationally efficient than our global SBMR method, especially on highly clustered computational grids in laminar and turbulent flow computations.

Choi, Kwang-Yoon↗

RANGE: A robust adaptive nature-inspired global explorer of potential energy surfaces

With the growing demand for realistic representations of chemical structures and the advent of exascale computing, the intelligent sampling of potential energy surfaces and efficient identification of global minima have become more essential but also more feasible. Building on prior studies demonstrating the efficiency of the Artificial Bee Colony (ABC) swarm intelligence algorithm, we report a hybrid metaheuristic framework that integrates the adaptive exploration capabilities of ABC coupled with the exploitation strengths of genetic algorithms (GA) in a scalable, Python-based implementation. The resulting tool, RANGE (Robust Adaptive Nature-inspired Global Explorer), provides seamless interfaces to multiple potential energy evaluators, either directly or via widely used Python libraries, and is designed for high-performance computing environments. We describe the implementation details of RANGE and evaluate its performance, relative to ABC- or GA-alone based algorithms, on a variety of chemical systems, including molecular clusters and heterogeneous surfaces. In conclusion, our results demonstrate RANGE’s efficiency, robustness, and broad applicability in addressing challenging global optimization problems in computational chemistry and materials science.

Algorithms and data structure↗

Stringent σ8 constraints from small-scale galaxy clustering using a hybrid MCMC + emulator framework

ABSTRACT We present a novel simulation-based hybrid emulator approach that maximally derives cosmological and Halo Occupation Distribution (HOD) information from non-linear galaxy clustering, with sufficient precision for DESI Year 1 (Y1) analysis. Our hybrid approach first samples the HOD space on a fixed cosmological simulation grid to constrain the high-likelihood region of cosmology + HOD parameter space, and then constructs the emulator within this constrained region. This approach significantly reduces the parameter volume emulated over, thus achieving much smaller emulator errors with fixed number of training points. We demonstrate that this combined with state-of-the-art simulations result in tight emulator errors comparable to expected DESI Y1 LRG sample variance. We leverage the new abacussummit simulations and apply our hybrid approach to CMASS non-linear galaxy clustering data. We infer constraints on σ8 = 0.762 ± 0.024 and fσ8(zeff = 0.52) = 0.444 ± 0.016, the tightest among contemporary galaxy clustering studies. We also demonstrate that our fσ8 constraint is robust against secondary biases and other HOD model choices, a critical first step towards showcasing the robust cosmology information accessible in non-linear scales. We speculate that the additional statistical power of DESI Y1 should tighten the growth rate constraints by at least another 50–60 ${{\ \rm per\ cent}}$, significantly elucidating any potential tension with Planck. We also address the ‘lensing is low’ tension, which we find to be in the same direction as a potential tension in fσ8. We show that the combined effect of a lower fσ8 and environment-based bias accounts for approximately $50{{\ \rm per\ cent}}$ of the discrepancy.

79 ASTRONOMY AND ASTROPHYSICS↗

Efficient Source of Shaped Single Photons Based on an Integrated Diamond Nanophotonic System

An efficient, scalable source of shaped single photons that can be directly integrated with optical fiber networks and quantum memories is at the heart of many protocols in quantum information science. We demonstrate a deterministic source of arbitrarily temporally shaped single-photon pulses with high efficiency [detection efficiency = 14.9 %] and purity [g (2) (0) = 0.0168] and streams of up to 11 consecutively detected single photons using a silicon-vacancy center in a highly directional fiber-integrated diamond nanophotonic cavity. Finally, combined with previously demonstrated spin-photon entangling gates, this system enables on-demand generation of streams of correlated photons such as cluster states and could be used as a resource for robust transmission and processing of quantum information.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

XRISM Reveals Low Nonthermal Pressure in the Core of the Hot, Relaxed Galaxy Cluster A2029

We present XRISM Resolve observations of the core of the hot, relaxed galaxy cluster Abell 2029 (A2029). We find that the line-of-sight bulk velocity of the intracluster medium (ICM) within the central 180 kpc is at rest with respect to the brightest cluster galaxy, with a 3σ upper limit of ∣vbulk∣ < 100 km s−1. We robustly measure the field-integrated ICM velocity dispersion to be σv = 169 ± 10 km s−1, obtaining similar results for both single-temperature and two-temperature plasma models to account for the cluster cool core. This result, if ascribed to isotropic turbulence, implies a subsonic ICM with Mach number M 3D ≈ 0.22 and a nonthermal pressure fraction of 2.6 ± 0.3%. The turbulent velocity is similar to what was measured in the core of the Perseus cluster by Hitomi, but here in a more massive cluster with an ICM temperature of 7 keV, the limit on the nonthermal pressure fraction is even more stringent. Our result is consistent with expectations from simulations of relaxed clusters, but it is on the low end of the predicted distribution, indicating that A2029 is an exceptionally relaxed cluster with no significant impacts from either a recent minor merger or active galactic nucleus activity.

Audard, Marc (ORCID:000000034721034X)↗

On the physical size of the Milky Way globular cluster NGC 7089 (M2)

ABSTRACT We study the outer regions of the Milky Way globular cluster NGC 7089 based on new Dark Energy Camera observations. The resulting background-cleaned stellar density profile reveals the existence of an extended envelope. We confirm previous results that cluster stars are found out up to ∼1° from the cluster’s centre, which is nearly three times the value of the most robust tidal radii estimations. We also used results from direct N-body simulations in order to compare with the observations. We found a fairly good agreement between the observed and numerically generated stellar density profiles. Because of the existence of gaps and substructures along globular cluster tidal tails, we closely examined the structure of the outer cluster region beyond the Jacobi radius. We extended the analysis to a sample of 35 globular clusters, 20 of them with observed tidal tails. We found that if the stellar density profile follows a power law ∝ r−α, the α slope correlates with the globular cluster present mass, in the sense that, the more massive the globular cluster, the smaller the α value. This trend is not found in globular clusters without observed tidal tails. The origin of such a phenomenon could be related, among other reasons, to the proposed so-called potential escapers or to the formation of globular clusters within dark matter minihaloes.

79 ASTRONOMY AND ASTROPHYSICS↗

The Dark Energy Survey Year 3 high-redshift sample: selection, characterization, and analysis of galaxy clustering

ABSTRACT The fiducial cosmological analyses of imaging surveys like DES typically probe the Universe at redshifts z < 1. We present the selection and characterization of high-redshift galaxy samples using DES Year 3 data, and the analysis of their galaxy clustering measurements. In particular, we use galaxies that are fainter than those used in the previous DES Year 3 analyses and a Bayesian redshift scheme to define three tomographic bins with mean redshifts around z ∼ 0.9, 1.2, and 1.5, which extend the redshift coverage of the fiducial DES Year 3 analysis. These samples contain a total of about 9 million galaxies, and their galaxy density is more than 2 times higher than those in the DES Year 3 fiducial case. We characterize the redshift uncertainties of the samples, including the usage of various spectroscopic and high-quality redshift samples, and we develop a machine-learning method to correct for correlations between galaxy density and survey observing conditions. The analysis of galaxy clustering measurements, with a total signal to noise S/N ∼ 70 after scale cuts, yields robust cosmological constraints on a combination of the fraction of matter in the Universe Ωm and the Hubble parameter h, $\Omega _m h = 0.195^{+0.023}_{-0.018}$, and 2–3 per cent measurements of the amplitude of the galaxy clustering signals, probing galaxy bias and the amplitude of matter fluctuations, bσ8. A companion paper (in preparation) will present the cross-correlations of these high-z samples with cosmic microwave background lensing from Planck and South Pole Telescope, and the cosmological analysis of those measurements in combination with the galaxy clustering presented in this work.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Dark energy survey year 3 results: cosmology from galaxy clustering and galaxy–galaxy lensing in harmonic space

We present the joint tomographic analysis of galaxy-galaxy lensing and galaxy clustering in harmonic space (HS), using galaxy catalogues from the first three years of observations by the Dark Energy Survey (DES Y3). We utilize the redMaGiC and MagLim catalogues as lens galaxies and the metacalibration catalogue as source galaxies. The measurements of angular power spectra are performed using the pseudo-$C_\ell$ method, and our theoretical modelling follows the fiducial analyses performed by DES Y3 in configuration space, accounting for galaxy bias, intrinsic alignments, magnification bias, shear magnification bias and photometric redshift uncertainties. We explore different approaches for scale cuts based on non-linear galaxy bias and baryonic effects contamination. Our fiducial covariance matrix is computed analytically, accounting for mask geometry in the Gaussian term, and including non-Gaussian contributions and super-sample covariance terms. To validate our HS pipelines and covariance matrix, we used a suite of 1800 log-normal simulations. We also perform a series of stress tests to gauge the robustness of our HS analysis. In the $\Lambda$CDM model, the clustering amplitude $S_8 =\sigma _8(\Omega _m/0.3)^{0.5}$ is constrained to $S_8 = 0.704\pm 0.029$ and $S_8 = 0.753\pm 0.024$ (68 per cent C.L.) for the redMaGiC and MagLim catalogues, respectively. For the wCDM, the dark energy equation of state is constrained to $w = -1.28 \pm 0.29$ and $w = -1.26^{+0.34}_{-0.27}$, for redMaGiC and MagLim catalogues, respectively. These results are compatible with the corresponding DES Y3 results in configuration space and pave the way for HS analyses using the DES Y6 data.

(cosmology:) cosmological parameters↗

Weak-lensing Detection of Intercluster Filaments in Three Nearby Cluster Systems

Abstract Direct detection of intercluster filaments is challenging due to their low surface density, resulting in a weak deflection field. We present weak-lensing detections of intercluster filaments using wide-field Dark Energy Camera observations from the Local Volume Complete Cluster Survey. A matched-filter method was applied to identify filamentary structures in three nearby ( z < 0.1) systems centered on A401, A2029, and A3558. We discover two filaments (>3 σ ) in each system, with the strongest detections (5.2 σ –5.8 σ ) around A401 and A2029. In particular, we report the first robust weak-lensing detections (≳5 σ ) of the intercluster bridges connecting the cluster pairs A401/399, A2029/2033, A2029/SIG, and A3558/3556. Adopting a filament convergence model motivated by numerical simulations, we infer the maximum convergence ( κ 0 ) and characteristic width ( h c ) for all six filaments, yielding κ 0 ∼ 0.016–0.040 and h c ∼ 0.23–0.43 Mpc. The performance of the matched-filter technique is validated using mock shear catalogs and further tested on a null field around A2351. We explore the potential of using the B-mode lensing signal of filaments to suppress cluster-induced shear contamination. We also quantify the biasing effect of closely separated terminal clusters to the filament signal. These results demonstrate the feasibility of directly mapping dark matter filaments with current and future wide-field weak-lensing datasets.

Shinde, Rahul [Brown University] (ORCID:0000000273↗