Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Spectral clustering”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

Efficient Network Partitioning: Application for Decentralized State Estimation in Power Distribution Grids: Preprint

Increase in the proliferation of DERs requires real-time situational awareness for efficient grid operations. State estimation plays an important role for real time control and management of the power grid. As the sensing infrastructure grows, aggregating and handling high volumes of data at a centralized location is extremely difficult. To address this challenge, this paper first proposes a novel and efficient hierarchical spectral clustering-based network partition algorithm followed by a decentralized compressive sensing (DCS) based state estimation. The applicability of the proposed network partitioning algorithm is tested on IEEE-123 bus, IEEE-8500 node, and a 6204-node distribution network. The results shows that the proposed approach efficiently divides the network into multiple sub-networks with the minimum edge connections among the neighbors. Then, we perform DCS-based state estimation on the 6204-node distribution network after dividing the network into 18 optimal partitions. Simulation results show that DCS-based state estimation recovers the system states with high accuracy and low complexity.

ADMM↗

Efficient Network Partitioning: Application for Decentralized State Estimation in Power Distribution Grids

Increase in the proliferation of distributed energy resources require real-time situational awareness for efficient grid operations. State estimation plays an important role for the real-time control and management of the power grid. As the sensing infrastructure grows, aggregating and handling high volumes of data at a centralized location is extremely difficult. To address this challenge, this paper first proposes a novel and efficient hier-archical spectral clustering-based network partitioning algorithm followed by a decentralized compressive sensing (DCS)-based state estimation. The applicability of the proposed network partitioning algorithm is tested on an IEEE 123-bus network, an IEEE 8,500-node system, and a 6,000+ node distribution network. The results shows that the proposed approach efficiently divides the network into multiple sub-networks with the minimum number of edge connections among the neighbors. Then, we perform DCS-based state estimation on the 6,000+ node distribution network after dividing the network into 18 optimal partitions. Simulation results show that the DCS-based state estimation recovers the system states with high accuracy and low complexity.

alternating direction method of multipliers↗

Mapping the Temperature-dependent and network site-specific onset of spectral diffusion at the surface of a water cluster cage

We explore the kinetic processes that sustain equilibrium in a microscopic, finite system. This is accomplished by monitoring the spontaneous, time-dependent frequency evolution (the frequency autocorrelation) of a single OH oscillator, embedded in a water cluster held in a temperature-controlled ion trap. The measurements are carried out by applying two-color, IR-IR photodissociation mass spectrometry to the D3O+?(HDO)(D2O)19 isotopologue of the “magic number” protonated water cluster, H+?(H2O)21. The OH group can occupy any one of the five spectroscopically distinct sites in the distorted pentagonal dodecahedron cage structure. The OH frequency is observed to evolve over tens of milliseconds in the temperature range (90-120 K). Starting at 100 K, large “jumps” are observed between two OH frequencies separated by ~300 cm-1 indicating migration of the OH group from the bound OH site at 3350 cm-1 to the free position at 3686 cm-1. Increasing the temperature to 110 K leads to partial interconversion among many sites. All sites are observed to interconvert at 120 K such that the distribution of the unique OH group among them adopts the form one would expect for a canonical ensemble. The spectral dynamics displayed by the clusters thus offer an unprecedented view into the molecular-level processes that drive spectral diffusion in an extended network of water molecules.

Yang, Nan↗

A robust clustering algorithm for analysis of composition-dependent organic aerosol thermal desorption measurements

Abstract. One of the challenges of understanding atmospheric organic aerosol (OA) particles stems from its complex composition. Mass spectrometry is commonly used to characterize the compositional variability of OA. Clustering of a mass spectral dataset helps identify components that exhibit similar behavior or have similar properties, facilitating understanding of sources and processes that govern compositional variability. Here, we developed an algorithm for clustering mass spectra, the noise-sorted scanning clustering (NSSC), appropriate for application to thermal desorption measurements of collected OA particles from the Filter Inlet for Gases and AEROsols coupled to a chemical ionization mass spectrometer (FIGAERO-CIMS). NSSC, which extends the common density-based special clustering of applications with noise (DBSCAN) algorithm, provides a robust, reproducible analysis of the FIGAERO temperature-dependent mass spectral data. The NSSC allows for the determination of thermal profiles for compositionally distinct clusters of mass spectra, increasing the accessibility and enhancing the interpretation of FIGAERO data. Applications of NSSC to several laboratory biogenic secondary organic aerosol (BSOA) systems demonstrate the ability of NSSC to distinguish different types of thermal behaviors for the components comprising the particles along with the relative mass contributions and chemical properties (e.g., average molecular formula) of each mass spectral cluster. For each of the systems examined, more than 80 % of the total mass is clustered into 9–13 mass spectral clusters. Comparison of the average thermograms of the mass spectral clusters between systems indicates some commonality in terms of the thermal properties of different BSOA, although with some system-specific behavior. Application of NSSC to sets of experiments in which one experimental parameter, such as the concentration of NO, is varied demonstrates the potential for mass spectral clustering to elucidate the chemical factors that drive changes in the thermal properties of OA particles. Further quantitative interpretation of the thermograms of the mass spectral clusters will allow for a more comprehensive understanding of the thermochemical properties of OA particles.

54 ENVIRONMENTAL SCIENCES↗

Far-ultraviolet radiation from disk globular clusters

IUE spectra obtained in a survey of the metal-rich disk system of globular clusters are presented. Significant FUV fluxes were detected in the 1200-2000-A short-wavelength (SWP) range of the IUE Observatory in several disk globular clusters. These clusters are the most metal-rich known to have an FUV flux. Three clusters show spectral energy distrbutions (SEDs) clearly rising at shorter wavelengths, not unlike the upturns observed in the bulges of metal-rich elliptical galaxies. Several others with weak SWP detections appear to have flat or uncertain spectral energy distributions. Blue stragglers provide a possible explanation for flux redder than 2000 A in clusters showing weaker flux in the SWP region, and with flat or declining SEDs.

Rich, R. M.↗

Fog Intermittency and Critical Behavior

The intermittency of fog occurrence (the switching between fog and no-fog) is a key stochastic feature that plays a role in its duration and the amount of moisture available. Here, fog intermittency is studied by using the visibility time series collected during the month of July 2022 on Sable Island, Canada. In addition to the visibility, time series of air relative humidity and turbulent kinetic energy, putative variables akin to the formation and breakup conditions of fog, respectively, are also analyzed in the same framework to establish links between fog intermittency and the underlying atmospheric variables. Intermittency in the time series is quantified with their binary telegraph approximations to isolate clustering behavior from amplitude variations. It is shown that relative humidity and turbulent kinetic energy bound many stochastic features of visibility, including its spectral exponent, clustering exponent, and the growth of its block entropy slope. Although not diagnostic, the visibility time series displays features consistent with Pomeau–Manneville Type-III intermittency in its quiescent phase duration PDF scaling (−3/2), power spectrum scaling (−1/2), and signal amplitude PDF scaling (−2). The binary fog time series exhibits properties of self-organized criticality in the relation between its power spectrum scaling and quiescent phase duration distribution.

54 ENVIRONMENTAL SCIENCES↗

Probabilistic cluster labeling of imagery data

The problem of obtaining the probabilities of class labels for the clusters using spectral and spatial information from a given set of labeled patterns and their neighbors is considered. A relationship is developed between class and clusters conditional densities in terms of probabilities of class labels for the clusters. Expressions are presented for updating the a posteriori probabilities of the classes of a pixel using information from its local neighborhood. Fixed-point iteration schemes are developed for obtaining the optimal probabilities of class labels for the clusters. These schemes utilize spatial information and also the probabilities of label imperfections. Experimental results from the processing of remotely sensed multispectral scanner imagery data are presented.

Chittineni, C. B.↗

Semi-Supervised Data Summarization: Using Spectral Libraries to Improve Hyperspectral Clustering

Hyperspectral imagers produce very large images, with each pixel recorded at hundreds or thousands of different wavelengths. The ability to automatically generate summaries of these data sets enables several important applications, such as quickly browsing through a large image repository or determining the best use of a limited bandwidth link (e.g., determining which images are most critical for full transmission). Clustering algorithms can be used to generate these summaries, but traditional clustering methods make decisions based only on the information contained in the data set. In contrast, we present a new method that additionally leverages existing spectral libraries to identify materials that are likely to be present in the image target area. We find that this approach simultaneously reduces runtime and produces summaries that are more relevant to science goals.

Wagstaff, K. L.↗

Spectral function for 4 He using the Chebyshev expansion in coupled-cluster theory

Here, we compute spectral function for 4 He by combining coupled-cluster theory with an expansion of integral transforms into Chebyshev polynomials. Our method allows us to estimate the uncertainty of spectral reconstruction. The properties of the Chebyshev polynomials make the procedure numerically stable and considerably lower in memory usage than the typically employed Lanczos algorithm. We benchmark our predictions with other calculations in the literature and with electron-scattering data in the quasi-elastic peak. The spectral function formalism allows one to extend ab initio lepton-nucleus cross sections into the relativistic regime. This makes it a promising tool for modeling this process at higher-energy transfers. The results we present open the door for studies of heavier nuclei, important for the neutrino oscillation programs.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Manual interpretation of LANDSAT data

The task of the analyst throughout LACIE phases 1 and 2 consisted of outlining representative areas (fields) for all spectral classes within a segment on the basis of their appearance on the LANDSAT image products, and then labeling the crop type (wheat/nonwheat) within the selected training areas. For LACIE phase 3, a procedure was developed and implemented which incorporated clustering for spectral class definition and training statistics generation. The only analyst's task was crop type identification. The logical processes involved in interpretation are described, rather than a step-by-step description of analyst procedures.

Hay, C. M.↗

X-ray spectra of clusters of galaxies

Out of the many individual cases of clusters with spectral complexity reported in the literature, attention is given to the spectral complexity of the most recent data obtained by the Einstein Observatory and an attempt is made to explain at least a portion of the data in terms of models of radiatively-regulated accretion in the central parts of the clusters and/or onto massive galaxies. After a review of the range of models proposed, with emphasis on predicted temperature distribution, it is noted that all accretion models have temperatures which decrease toward the center, and that several authors have explored the consequences of radiative cooling of the gas in the dense central regions of a cluster or of a cluster plus a central galaxy.

Canizares, C. R.↗

Impact of environmental oxygen on nanoparticle formation and agglomeration in aluminum laser ablation plumes

Here, the role of ambient oxygen gas (O 2 ) on molecular and nanoparticle formation and agglomeration was studied in laser ablation plumes. As a lab-scale surrogate to a high explosion detonation event, nanosecond laser ablation of an aluminum alloy (AA6061) target was performed in atmospheric pressure conditions. Optical emission spectroscopy and two mass spectrometry techniques were used to monitor the early to late stages of plasma generation to track the evolution of atoms, molecules, clusters, nanoparticles, and agglomerates. The experiments were performed under atmospheric pressure air, atmospheric pressure nitrogen, and 20% and 5% O 2 (balance N 2 ), the latter specifically with in situ mass spectrometry. Electron microscopy was performed ex situ to identify crystal structure and elemental distributions in individual nanoparticles. We find that the presence of ≈20% O 2 leads to strong AlO emission, whereas in a flowing N 2 environment (with trace O 2 ), AlN and strong, unreacted Al emissions are present. In situ mass spectrometry reveals that as O 2 availability increases, Al oxide cluster size increases. Nanoparticle agglomerates formed in air are found to be larger than those formed under N 2 gas. High-resolution transmission electron microscopy demonstrates that Al 2 O 3 and AlN nanoparticle agglomerates are formed in both environments; indicating that the presence of trace O 2 can lead to Al 2 O 3 nanoparticle formation. The present results highlight that the availability of O 2 in the ambient gas significantly impacts spectral signatures, cluster size, and nanoparticle agglomeration behavior. These results are relevant to understanding debris formation in an explosion event, and interpreting data from forensic investigations.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Data-driven surrogates for high dimensional models using Gaussian process regression on the Grassmann manifold

This paper introduces a surrogate modeling scheme based on Grassmannian manifold learning to be used for cost-efficient predictions of high-dimensional stochastic systems. The method exploits subspace-structured features of each solution by projecting it onto a Grassmann manifold. This point-wise linear dimensionality reduction harnesses the structural information to assess the similarity between solutions at different points in the input parameter space. The method utilizes a solution clustering approach in order to identify regions of the parameter space over which solutions are sufficiently similarly such that they can be interpolated on the Grassmannian. In this clustering, the reduced-order solutions are partitioned into disjoint clusters on the Grassmann manifold using the eigen-structure of properly defined Grassmannian kernels and, the Karcher mean of each cluster is estimated. Then, the points in each cluster are projected onto the tangent space with origin at the corresponding Karcher mean using the exponential mapping. For each cluster, a Gaussian process regression model is trained that maps the input parameters of the system to the reduced solution points of the corresponding cluster projected onto the tangent space. Using this Gaussian process model, the full-field solution can be efficiently predicted at any new point in the parameter space. In certain cases, the solution clusters will span disjoint regions of the parameter space. In such cases, for each of the solution clusters we utilize a second, density-based spatial clustering to group their corresponding input parameter points in the Euclidean space. The proposed method is applied to two numerical examples. Here, the first is a nonlinear stochastic ordinary differential equation with uncertain initial conditions where the surrogate is used to predict the time history solution. The second involves modeling of plastic deformation in a model amorphous solid using the Shear Transformation Zone theory of plasticity, where the proposed surrogate is used to predict the full strain field of a material specimen under large shear strains.

42 ENGINEERING↗

GMFOLD: Subgraph matching for high-throughput DNA-aptamer secondary structure classification and machine learning interpretability

Aptamers are oligonucleotide receptors that bind to their targets with high affinity. Here, we consider aptamers comprised of single-stranded DNA that undergo target-binding-induced conformational changes, giving rise to unique secondary and tertiary structures. Given a specific aptamer primary sequence, there are well-established computational tools (notably mfold) to predict the secondary structure via free energy minimization algorithms. While mfold generates secondary structures for individual sequences, there is a need for a high-throughput process whereby thousands of DNA structures can be predicted in real-time for use in an interactive setting, when combined with aptamer selections that generate candidate pools that are too large to be experimentally interrogated. We developed a new Python code for high-throughput aptamer secondary structure determination (GMfold). GMfold uses subgraph matching methods to group aptamer candidates by secondary structure similarities. We also improve an open-source code, SeqFold, to incorporate subgraph matching concepts. We represent each secondary structure as a lowest-energy bipartite subgraph matching of the DNA graph to itself. These new tools enable thousands of DNA sequences to be compared based on their secondary structures, using machine-learning algorithms. This process is advantageous when analyzing sequences that arise from aptamer selections via systematic evolution of ligands by exponential enrichment (SELEX). This work is a building block for future machine-learning-informed DNA-aptamer selection processes to identify aptamers with improved target affinity and selectivity and advance aptamer biosensors and therapeutics.

Aptamer↗

Adaptive Hierarchical Cyber Attack Detection and Localization in Active Distribution Systems

Development of a cyber security strategy for the active distribution systems is challenging due to the inclusion of distributed renewable energy generations. Here this paper proposes an adaptive hierarchical cyber attack detection and localization framework for distributed active distribution systems via analyzing electrical waveforms. Cyber attack detection is based on a sequential deep learning model, via which even minor cyber attacks can be identified. The two-stage cyber attack localization algorithm first estimates the cyber attack sub-region, and then localize the specified cyber attack within the estimated subregion. We propose a modified spectral clustering-based network partitioning method for the hierarchical cyber attack ‘coarse’ localization. Next, to further narrow down the cyber attack location, a normalized impact score based on waveform statistical metrics is proposed to obtain a ‘fine’ cyber attack location by characterizing different waveform properties. Finally, compared with classical and state-of-art methods, a comprehensive quantitative evaluation with two case studies shows promising estimation results of the proposed framework.

42 ENGINEERING↗

Multivariate Analysis of Sensor Data NSARD Poster

Poster for the cancelled NSARD meeting. The Federal Program Manager, Angie Waterworth, plans to send posters out to all planned attendees, even though the poster will NO LONGER BE PRESENTED.

97 MATHEMATICS AND COMPUTING↗

Probing U 5f Covalency in Uranium Compounds through Oxidant 2p Bonding

Oxidant K (1s) X-ray Emission Spectroscopy (XES) was used to investigate covalency in the Oxidant 2p – Uranium 5f bonds of Uranium Dioxide (UO 2 ) and Uranium Tetrafluoride (UF 4 ). It will be shown that the width of the 2p Occupied Density of States (ODOS), determined from the XES measurements, correlates with the increased 5f covalency of UO 2 and the increased ionicity of UF 4 . Analysis of the XES results includes comparison to spectral simulations, cluster calculations and peak fitting, demonstrating the potential of the oxidant XES measurements as a probe of 5f covalency.

38 RADIATION CHEMISTRY, RADIOCHEMISTRY, AND NUCLEA↗

Observations of faint field galaxies

Number counts, colors, and angular correlations of field galaxies fainter than 20th mag are summarized. Resulting conclusions regarding the presence and nature of luminosity, spectral, and clustering evolution remain contraversial. Preliminary analysis of two major spectroscopic surveys near completion suggests that by z approximately 0.5, larger numbers of very blue galaxies of moderate luminosities are found than today. The skewer-like surveys also provide new probes of galaxy clustering on scales previously unexplored (larger than 200 Mpc) and over lookback times of several billion years.

Koo, David C.↗