Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “multivariate data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 145 records · Page 8

Advanced Interactive 3D Visualization Tool for Customizable Analyses of Tomography Datasets in Material Science

Current methods for visualizing and analyzing 3D tomography datasets in materials science often lack the interactivity and depth required for detailed structural insights. This limitation restricts a researchers' ability to accurately interpret complex data, which is critical for advancing material innovations and understanding structural properties. To address this issue, we have developed a novel, web-based interactive 3D visualization and analysis tool from the Trame framework that offers customizable features to enhance data interpretability. The tool allows users to adjust parameters such as visible range, slice planes, data rotation, and layering, providing a more detailed and dynamic view of complex structures. Its user-friendly web interface increases the accessibility and ease of use for both novice and experienced researchers, to visualize large volumetric datasets. The tool supports a diverse range of data formats, making it versatile for various research applications. Unique capabilities include real-time data manipulation, automated feature detection, context-sensitive feedback, and real-time volume calculations and distributions per sliced region or layer, alongside the ability to quickly generate high-quality screenshots and videos for presentations and reports. These advancements offer a comprehensive solution for enhanced 3D data exploration, significantly improving the analysis process and communication of results in materials science.

36 - MATERIALS SCIENCE↗

Digital preprocessing and classification of multispectral earth observation data

The development of airborne and satellite multispectral image scanning sensors has generated wide-spread interest in application of these sensors to earth resource mapping. These point scanning sensors permit scenes to be imaged in a large number of electromagnetic energy bands between .3 and 15 micrometers. The energy sensed in each band can be used as a feature in a computer based multi-dimensional pattern recognition process to aid in interpreting the nature of elements in the scene. Images from each band can also be interpreted visually. Visual interpretation of five or ten multispectral images simultaneously becomes impractical especially as area studied increases; hence, great emphasis has been placed on machine (computer) techniques for aiding in the interpretation process. This paper describes a computer software system concept called LARSYS for analysis of multivariate image data and presents some examples of its application.

Anuta, P. E.↗

Skin Temperature Analysis and Bias Correction in a Coupled Land-Atmosphere Data Assimilation System

In an initial investigation, remotely sensed surface temperature is assimilated into a coupled atmosphere/land global data assimilation system, with explicit accounting for biases in the model state. In this scheme, an incremental bias correction term is introduced in the model's surface energy budget. In its simplest form, the algorithm estimates and corrects a constant time mean bias for each gridpoint; additional benefits are attained with a refined version of the algorithm which allows for a correction of the mean diurnal cycle. The method is validated against the assimilated observations, as well as independent near-surface air temperature observations. In many regions, not accounting for the diurnal cycle of bias caused degradation of the diurnal amplitude of background model air temperature. Energy fluxes collected through the Coordinated Enhanced Observing Period (CEOP) are used to more closely inspect the surface energy budget. In general, sensible heat flux is improved with the surface temperature assimilation, and two stations show a reduction of bias by as much as 30 Wm(sup -2) Rondonia station in Amazonia, the Bowen ratio changes direction in an improvement related to the temperature assimilation. However, at many stations the monthly latent heat flux bias is slightly increased. These results show the impact of univariate assimilation of surface temperature observations on the surface energy budget, and suggest the need for multivariate land data assimilation. The results also show the need for independent validation data, especially flux stations in varied climate regimes.

Bosilovich, Michael G.↗

Toward autonomous design and synthesis of novel inorganic materials

Autonomous experimentation driven by artificial intelligence (AI) provides an exciting opportunity to revolutionize inorganic materials discovery and development. Herein, we review recent progress in the design of self-driving laboratories, including robotics to automate materials synthesis and characterization, in conjunction with AI to interpret experimental outcomes and propose new experimental procedures. We focus on efforts to automate inorganic synthesis through solution-based routes, solid-state reactions, and thin film deposition. In each case, connections are made to relevant work in organic chemistry, where automation is more common. Characterization techniques are primarily discussed in the context of phase identification, as this task is critical to understand what products have formed during synthesis. The application of deep learning to analyze multivariate characterization data and perform phase identification is examined. To achieve “closed-loop” materials synthesis and design, we further provide a detailed overview of optimization algorithms that use active learning to rationally guide experimental iterations. Lastly, we highlight several key opportunities and challenges for the future development of self-driving inorganic materials synthesis platforms.

36 MATERIALS SCIENCE↗

CHMMPY: A python package for constrained Hidden Markov Models

SAND2025-11909O chmmpy software analyzes multivariate timeseries data to detect patterns. It uses a Hidden Markov Model (HMM) and application-specific constraints that reflect known relationships among hidden states to accomplish this. The chmmpy software provides a generic framework for expressing application-specific constraints and supporting constrained HMM inference using optimization solvers. chmmpy is available on GitHub. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Hart, William↗

Symposium on Machine Processing of Remotely Sensed Data, Purdue University, West Lafayette, Ind., June 29-July 1, 1976, Proceedings

Papers are presented on the applicability of Landsat data to water management and control needs, IBIS, a geographic information system based on digital image processing and image raster datatype, and the Image Data Access Method (IDAM) for the Earth Resources Interactive Processing System. Attention is also given to the Prototype Classification and Mensuration System (PROCAMS) applied to agricultural data, the use of Landsat for water quality monitoring in North Carolina, and the analysis of geophysical remote sensing data using multivariate pattern recognition. The Illinois crop-acreage estimation experiment, the Pacific Northwest Resources Inventory Demonstration, and the effects of spatial misregistration on multispectral recognition are also considered. Individual items are announced in this issue.

Source record↗

The Tropical Convective Spectrum: Archetypal Vertical Structures - Part 1

A taxonomy of tropical convective and stratiform vertical structures is constructed through cluster analysis of 3 yr of Tropical Rainfall Measuring Mission (TRMM) "warm-season" (surface temperature greater than 10 C) precipitation radar (PR) vertical profiles, their surface rainfall, and associated radar-based classifiers (convective/ stratiform and brightband existence). Twenty-five archetypal profile types are identified, including nine convective types, eight stratiform types, two mixed types, and six anvil/fragment types (nonprecipitating anvils and sheared deep convective profiles). These profile types are then hierarchically clustered into 10 similar families, which can be further combined, providing an objective and physical reduction of the highly multivariate PR data space that retains vertical structure information. The taxonomy allows for description of any storm or local convective spectrum by the profile types or families. The analysis provides a quasi-independent corroboration of the TRMM 2A23 convective/ stratiform classification. The global frequency of occurrence and contribution to rainfall for the profile types are presented, demonstrating primary rainfall contribution by midlevel glaciated convection (27%) and similar depth decaying/stratiform stages (28%-31%). Profiles of these types exhibit similar 37- and 85-GHz passive microwave brightness temperatures but differ greatly in their frequency of occurrence and mean rain rates, underscoring the importance to passive microwave rain retrieval of convective/stratiform discrimination by other means, such as polarization or texture techniques, or incorporation of lightning observations. Close correspondence is found between deep convective profile frequency and annualized lightning production, and pixel-level lightning occurrence likelihood directly tracks the estimated mean ice water path within profile types.

Precipitation radar↗

System and Method for Outlier Detection via Estimating Clusters

An efficient method and system for real-time or offline analysis of multivariate sensor data for use in anomaly detection, fault detection, and system health monitoring is provided. Models automatically derived from training data, typically nominal system data acquired from sensors in normally operating conditions or from detailed simulations, are used to identify unusual, out of family data samples (outliers) that indicate possible system failure or degradation. Outliers are determined through analyzing a degree of deviation of current system behavior from the models formed from the nominal system data. The deviation of current system behavior is presented as an easy to interpret numerical score along with a measure of the relative contribution of each system parameter to any off-nominal deviation. The techniques described herein may also be used to "clean" the training data.

Iverson, David J.↗

Molybdenum Valence in Basaltic Silicate Melts

The moderately siderophile element molybdenum has been used as an indicator in planetary differentiation processes, and is particularly relevant to core formation [for example, 1-6]. However, models that apply experimental data to an equilibrium differentiation scenario infer the oxidation state of molybdenum from solubility data or from multivariable coefficients from metal-silicate partitioning data [1,3,7]. Partitioning behavior of molybdenum, a multivalent element with a transition near the J02 of interest for core formation (~IW-2) will be sensitive to changes in JO2 of the system and silicate melt structure. In a silicate melt, Mo can occur in either 4+ or 6+ valence state, and Mo6+ can be either octahedrally or tetrahedrally coordinated. Here we present first XANES measurements of Mo valence in basaltic run products at a range of P, T, and JO2 and further quantify the valence transition of Mo.

Danielson, L. R.↗

Molybdenum Valence in Basaltic Silicate Melts: Effects of Temperature and Pressure

The metal-silicate partitioning behavior of molybdenum has been used as a test for equilibrium core formation hypotheses [for example, 1-6]. However, current models that apply experimental data to equilibrium core-mantle differentiation infer the oxidation state of molybdenum from solubility data or from multivariable coefficients from metal-silicate partitioning data [1,3,7]. Molybdenum, a multi-valent element with a valence transition near the fO2 of interest for core formation (approx.IW-2) will be sensitive to changes in fO2 of the system and silicate melt structure. In a silicate melt, Mo can occur in either 4+ or 6+ valence state, and Mo(6+) can be either octahedrally or tetrahedrally coordinated. Here we present X-ray absorption near edge structure (XANES) measurements of Mo valence in basaltic run products at a range of P, T, and fO2 and further quantify the valence transition of Mo.

Danielson, L. R.↗

Identification of linear multivariable systems from a single set of data by identification of observers with assigned real eigenvalues

A formulation is presented for identification of linear multivariable from a single set of input-output data. The identification method is formulated with the mathematical framework of learning identifications, by extension of the repetition domain concept to include shifting time intervals. This method contrasts with existing learning approaches that require data from multiple experiments. In this method, the system input-output relationship is expressed in terms of an observer, which is made asymptotically stable by an embedded real eigenvalue assignment procedure. Through this relationship, the Markov parameters of the observer are identified. The Markov parameters of the actual system are recovered from those of the observer, and then used to obtain a state space model of the system by standard realization techniques. The basic mathematical formulation is derived, and numerical examples presented to illustrate.

Phan, Minh↗

Identification of linear multivariable systems from a single set of data by identification of observers with assigned real eigenvalues

This paper presents a formulation for identification of linear multivariable systems from a single set of input-output data. The identification method is formulated with the mathematical framework of learning identification, by extension of the repetition domain concept to include shifting time intervals. This contrasts existing learning approaches that require data from multiple experiments. In this method, the system input-output relationship is expressed in terms of an observer, which is made asymptotically stable by an embedded real eigenvalue assignment procedure. Through this relationship, the Markov parameters of the observer are identified. The Markov parameters of the actual system are recovered from those of the observer, and then used to obtain a state space model of the system by standard realization techniques. The basic mathematical formulation is derived, and numerical examples presented to illustrate the proposed method.

Phan, Minh↗

A Deterministic Multivariate Clustering Method for Drive Cycle Generation from In-Use Vehicle Data

Accurately characterizing vehicle drive cycles plays a fundamental role in assessing the performance of new vehicle technologies. Repeatable, short duration representative drive cycles facilitate more informed decision making, resulting in improved test procedures and more successful vehicle designs. With continued growth in the deployment of onboard telematics systems employing global positioning systems (GPS), large scale, low cost collection of real-world vehicle drive cycle data has become a reality. As a result of these technological advances, researchers, designers, and engineers are no longer constrained by lack of operating data when developing and optimizing technology, but rather by resources available for testing and simulation. Experimental testing is expensive and time consuming, therefore the need exists for a fast and accurate means of generating representative cycles from large volumes of real-world driving data. This paper explores the development and initial validation of a method of generating representative drive cycles from large collections of real-world vehicle data using a deterministic multivariate clustering approach. Starting with theory and diving into the methodology behind representative cycle generation, the paper aims to also present graphical and tabular results of initial validation via vehicle simulation and chassis dynamometer testing. Additional topics for further research and areas for ongoing development will also be presented.

47 OTHER INSTRUMENTATION↗

Multidimensional scaling informed by F -statistic: Visualizing grouped microbiome data with inference

Multidimensional scaling (MDS) is a widely used dimensionality reduction technique in microbial ecology data analysis that captures the multivariate structure of the data while preserving pairwise distances between samples. While improvements in MDS have enhanced the ability to reveal group-specific data patterns, these MDS-based methods require prior assumptions for inference, limiting their application in general microbiome analysis. Here, in this study, we introduce a new MDS-based ordination method, “F-informed MDS,” which configures the data distribution based on the F-statistic, the ratio of dispersion between groups sharing common and different characteristics. Using semisynthetic datasets, we demonstrate that the proposed method is robust to hyperparameter selection while maintaining statistical significance throughout the ordination process. Various quality metrics for evaluating dimensionality reduction confirm that F-informed MDS is comparable to state-of-the-art methods in preserving both local and global data structures. Its application to a diatom-associated bacterial community suggests the role of this new method in interpreting the community’s response to the host. Our approach offers a well-founded refinement of MDS that aligns with statistical test results, which can be beneficial for broader multidimensional data analyses in microbiology and ecology. This new visualization tool can be incorporated into standard microbiome data analyses.

Biological and medical sciences↗

Materials surface contamination analysis

The original research objective was to demonstrate the ability of optical fiber spectrometry to determine contamination levels on solid rocket motor cases in order to identify surface conditions which may result in poor bonds during production. The capability of using the spectral features to identify contaminants with other sensors which might only indicate a potential contamination level provides a real enhancement to current inspection systems such as Optical Stimulated Electron Emission (OSEE). The optical fiber probe can easily fit into the same scanning fixtures as the OSEE. The initial data obtained using the Guided Wave Model 260 spectrophotometer was primarily focused on determining spectra of potential contaminants such as HD2 grease, silicones, etc. However, once we began taking data and applying multivariate analysis techniques, using a program that can handle very large data sets, i.e., Unscrambler 2, it became apparent that the techniques also might provide a nice scientific tool for determining oxidation and chemisorption rates under controlled conditions. As the ultimate power of the technique became recognized, considering that the chemical system which was most frequently studied in this work is water + D6AC steel, we became very interested in trying the spectroscopic techniques to solve a broad range of problems. The complexity of the observed spectra for the D6AC + water system is due to overlaps between the water peaks, the resulting chemisorbed species, and products of reaction which also contain OH stretching bands. Unscrambling these spectral features, without knowledge of the specific species involved, has proven to be a formidable task.

Workman, Gary L.↗

A CLIPS expert system for clinical flow cytometry data analysis

An expert system is being developed using CLIPS to assist clinicians in the analysis of multivariate flow cytometry data from cancer patients. Cluster analysis is used to find subpopulations representing various cell types in multiple datasets each consisting of four to five measurements on each of 5000 cells. CLIPS facts are derived from results of the clustering. CLIPS rules are based on the expertise of Drs. Stewart, Duque, and Braylan. The rules incorporate certainty factors based on case histories.

Salzman, G. C.↗

Multiple-Flat-Panel System Displays Multidimensional Data

The NASA Ames hyperwall is a display system designed to facilitate the visualization of sets of multivariate and multidimensional data like those generated in complex engineering and scientific computations. The hyperwall includes a 77 matrix of computer-driven flat-panel video display units, each presenting an image of 1,280 1,024 pixels. The term hyperwall reflects the fact that this system is a more capable successor to prior computer-driven multiple-flat-panel display systems known by names that include the generic term powerwall and the trade names PowerWall and Powerwall. Each of the 49 flat-panel displays is driven by a rack-mounted, dual-central-processing- unit, workstation-class personal computer equipped with a hig-hperformance graphical-display circuit card and with a hard-disk drive having a storage capacity of 100 GB. Each such computer is a slave node in a master/ slave computing/data-communication system (see Figure 1). The computer that acts as the master node is similar to the slave-node computers, except that it runs the master portion of the system software and is equipped with a keyboard and mouse for control by a human operator. The system utilizes commercially available master/slave software along with custom software that enables the human controller to interact simultaneously with any number of selected slave nodes. In a powerwall, a single rendering task is spread across multiple processors and then the multiple outputs are tiled into one seamless super-display. It must be noted that the hyperwall concept subsumes the powerwall concept in that a single scene could be rendered as a mosaic image on the hyperwall. However, the hyperwall offers a wider set of capabilities to serve a different purpose: The hyperwall concept is one of (1) simultaneously displaying multiple different but related images, and (2) providing means for composing and controlling such sets of images. In place of elaborate software or hardware crossbar switches, the hyperwall concept substitutes reliance on the human visual system for integration, synthesis, and discrimination of patterns in complex and high-dimensional data spaces represented by the multiple displayed images. The variety of multidimensional data sets that can be displayed on the hyperwall is practically unlimited. For example, Figure 2 shows a hyperwall display of surface pressures and streamlines from a computational simulation of airflow about an aerospacecraft at various Mach numbers and angles of attack. In this display, Mach numbers increase from left to right and angles of attack increase from bottom to top. That is, all images in the same column represent simulations at the same Mach number, while all images in the same row represent simulations at the same angle of attack. The same viewing transformations and the same mapping from surface pressure to colors were used in generating all the images.

Gundo, Daniel↗

Studies in Astronomical Time Series Analysis. VI. Bayesian Block Representations

This paper addresses the problem of detecting and characterizing local variability in time series and other forms of sequential data. The goal is to identify and characterize statistically significant variations, at the same time suppressing the inevitable corrupting observational errors. We present a simple nonparametric modeling technique and an algorithm implementing it-an improved and generalized version of Bayesian Blocks [Scargle 1998]-that finds the optimal segmentation of the data in the observation interval. The structure of the algorithm allows it to be used in either a real-time trigger mode, or a retrospective mode. Maximum likelihood or marginal posterior functions to measure model fitness are presented for events, binned counts, and measurements at arbitrary times with known error distributions. Problems addressed include those connected with data gaps, variable exposure, extension to piece- wise linear and piecewise exponential representations, multivariate time series data, analysis of variance, data on the circle, other data modes, and dispersed data. Simulations provide evidence that the detection efficiency for weak signals is close to a theoretical asymptotic limit derived by [Arias-Castro, Donoho and Huo 2003]. In the spirit of Reproducible Research [Donoho et al. (2008)] all of the code and data necessary to reproduce all of the figures in this paper are included as auxiliary material.

signal detection↗