Engineering PapersSearch

SEARCH · Engineering Papers

Results for “multimode”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Toward Intelligent Multimodal Holography for Real-Time Chemical Imaging of Dynamic Ion Separation

Molecular-level visualization of ion transport and separation dynamics in complex environments is crucial for advancing energy systems, water purification, and critical materials recovery. Achieving this requires imaging platforms that combine structural sensitivity, chemical specificity, and real-time operation. Digital off-axis holography (DOAH) provides high-throughput, label-free quantitative phase imaging but inherently lacks chemical selectivity. Integrating DOAH with complementary spectroscopic channels such as fluorescence or hyperspectral imaging introduces the needed molecular specificity, while also creating challenges in multimodal data fusion, synchronization, and computational throughput. Artificial intelligence offers a powerful route to address these limitations by uniting physics-based reconstruction with data-driven interpretation. In this Perspective, we outline a framework for intelligent multimodal holography and demonstrate its potential using a preliminary AI-driven test case. Raw DOAH holograms of lanthanide solutions subjected to magnetic field gradients were analyzed using multi-agent AI workflows that autonomously selected reconstruction tools, extracted NMF components, and generated scientific claims consistent with true paramagnetic and diamagnetic behavior. This demonstration shows how AI-enabled reasoning can deliver real-time chemical–structural interpretation directly from raw holograms. Together, these advances define a path toward adaptive, intelligent holography platforms capable of supporting in situ chemical separations, dynamic ion transport analysis, and next-generation interfacial science.

Ricchiuti, Giovanna

The POINTER Imaging baseline cohort: Associations between multimodal neuroimaging biomarkers, cardiovascular health, and cognition

Abstract INTRODUCTION The U.S. Study to Protect Brain Health Through Lifestyle Intervention to Reduce Risk (U.S. POINTER) is evaluating lifestyle interventions in older adults at risk for cognitive decline and dementia. Here we characterize the baseline data set of the POINTER Imaging ancillary study. METHODS Participants underwent health and cognitive assessments and neuroimaging with multimodal positron emission tomography (PET) (beta‐amyloid [Aβ] and tau) and magnetic resonance imaging (MRI). Framingham risk score (FRS) was used to quantify cardiovascular disease (CVD) risk. RESULTS A total of 1052 participants (31% from underrepresented ethnoracial groups) were enrolled. Compared to Aβ−, Aβ+ (29%) participants were older, had higher apolipoprotein E (APOE) ε4 carriage rate and white matter hyperintensity volume, and greater temporal tau. FRS was related to MRI measures, but not AD biomarkers. FRS and tau had independent effects on cognition. DISCUSSION In this heterogenous, at‐risk cohort, CVD risk was related to more abnormal brain structure and poorer cognition, representing a putative non‐AD (Alzheimer's disease) pathway to brain injury and cognitive decline. Highlights The U.S. Study to Protect Brain Health Through Lifestyle Intervention to Reduce Risk (U.S. POINTER) cohort is enriched for cardiovascular disease (CVD) and poor lifestyle POINTER Imaging collected multimodal neuroimaging data in this unique, at‐risk cohort Amyloid burden was related to age, apolipoprotein E (APOE) ε4 carriage, and measures of disease progression Associations between amyloid and tau, and tau and cognition, were relatively weak CVD risk and tau pathology were independently related to memory

Neurosciences & Neurology

Achieving Multimodal and Multicolor Luminescence in LaAlO 3 :Pr 3+ , Gd 3+ via Trap Engineering and Energy Transfer

Achieving multimodal luminescence within a single phosphor is vital for multifunctional applications but remains challenging due to complex color tuning and trap engineering. In this study, we report Pr 3+ and Gd 3+ co‐doped LaAlO 3 (LAO:PG) phosphors, designed through careful modulation of multilevel traps and Pr 3+ → Gd 3+ energy transfer dynamics. These materials exhibit diverse luminescence modes, including down‐conversion luminescence (DCL), up‐conversion luminescence (UCL), persistent luminescence (PersL), optically stimulated luminescence (OSL), and thermally stimulated luminescence (TSL) across a wide spectral range. Unlike previously studied Pr 3+ ‐doped LAO, the co‐doped LAO:PG shows DCL in both UV‐visible and NIR regions and displays ultraviolet‐C UCL under visible excitation. Notably, we observe, for the first time, PersL lasting several minutes in these phosphors—an improvement over the non‐PersL behavior of Pr 3+ ‐only doped LAO. Additionally, the LAO:PG phosphors exhibit strong OSL response. TSL analysis reveals five distinct trap levels linked to these properties. Density functional theory calculations further correlate intrinsic defects to these traps, supporting a proposed mechanism for the observed multimodal luminescence. These findings highlight LAO:PG as a promising platform for developing advanced phosphors with integrated luminescence modes, paving the way for future applications in data storage, phototherapy, and anti‐counterfeiting technologies.

Chemistry

Multimodal, microspectroscopic speciation of legacy phosphorus in two US mid-Atlantic agricultural soils

To understand phosphorus (P) mobility in agricultural soils and its potential environmental risk, it is essential to directly measure solid phase P speciation. Often, bulk P K-edge X-ray absorption near edge structure (XANES) spectroscopy followed by linear combination fitting (LCF) is utilized to determine the solid P phases in soil. However, this method may limit results to only a few major phases. Additionally, XANES spectra for different P species may have very similar features, leading to an over- or underestimate of their contribution to LCF. Here, an improved P speciation by pairing multimodal microbeam-X-ray fluorescence (µ-XRF) mapping coupled with µ-XANES (microbeam-X-ray absorption near edge structure) analysis to directly speciate major and minor P phases on the micron scale is provided. We combined maps of both tender (P, sulfur, aluminum, and silicon) and hard energy (calcium, iron [Fe], and manganese) elements to evaluate the elemental co-locations with P. To better account for uncertainty assigning XANES peaks to individual compounds, a more quantitative fingerprinting by “spectral feature analysis” was completed. With this analysis, an R-factor is reported for the fit. These results were compared to traditional LCF. Pre-edge fitting results revealed the presence of a two-component pre-edge feature for phosphate adsorbed to ferrihydrite. Additionally, phytate co-precipitated with ferrihydrite (Phytate-Fe-Cop) had a pre-edge feature, indicating direct association with Fe. Lastly, a unique P species associated with manganese oxide was identified in the soil via multimodal mapping and µ-XANES. These results allow for better prediction of P dissolution and mobility.

36 MATERIALS SCIENCE

Multimodal Nanoscale Mapping of Local Structure and CO 2 Adsorption in Metal–Organic Frameworks

Diamine functionalization of the metal−organic framework Mg 2 (dobpdc) (dobpdc 4− = 4,4′-dioxidobiphenyl-3,3′-dicarboxylate) significantly enhances its selectivity for CO 2 capture from flue gases and air. The structure and CO 2 capacity of such materials are typically assessed using bulk techniques that rely on averaging signal over large ensembles of unit cells, obscuring local heterogeneities, such as variations in CO 2 occupancy across individual nanocrystals. To resolve this limitation, we demonstrate a multimodal, nanoscale characterization of Mg 2 (dobpdc) appended with 1,3-diaminopropane. By employing recently developed characterization techniques at progressively smaller length scales, we uncover insights from correspondingly smaller populations of unit cells. First, we use parallel-beam 3D electron diffraction (3D ED) to identify a prominent expansion in lattice parameters upon desorption of CO 2 , as observed at the level of single nanocrystals. Second, we use convergent-probe 4D scanning transmission electron microscopy (4D-STEM) to quantify associated differences in lattice strain as a function of gas loading and diamine appending. These measurements sample small subvolumes within individual nanocrystals. Finally, we apply infrared scattering scanning near-field optical microscopy (IR s- SNOM) to confirm variable CO 2 chemisorption across adsorption sites at the surface of single nanocrystals. This multimodal, multiscale approach allows us to map heterogeneity within individual nanocrystals. Collectively, these findings emphasize the importance of local, nanoscale characterization of metal−organic frameworks in revealing previously unresolvable features that impact their performance.

Karstens, Sarah L. [University of California, Berk

Photon–photon chemical thermodynamics of frequency conversion processes in highly multimode systems

Abstract Frequency generation in highly multimode nonlinear optical systems is inherently a complex process, giving rise to an exceedingly convoluted landscape of evolution dynamics. While predicting and controlling the global conversion efficiencies in such nonlinear environments has long been considered impossible, here, we formally address this challenge even in scenarios involving a very large number of spatial modes. By utilizing fundamental notions from optical statistical mechanics, we develop a universal theoretical framework that effectively treats all frequency components as chemical reactants/products, capable of undergoing optical thermodynamic reactions facilitated by a variety of multi-wave mixing effects. These photon–photon reactions are governed by conservation laws that directly determine the optical temperatures and chemical potentials of the ensued chemical equilibria for each frequency species. In this context, we develop a comprehensive stoichiometric model and formally derive an expression that relates the chemical potentials to the optical stoichiometric coefficients, in a manner akin to atomic/molecular chemical reactions. This advancement unlocks new predictive capabilities that can facilitate the optimization of frequency generation in highly multimode photonic arrangements, surpassing the limitations of conventional schemes that rely exclusively on nonlinear optical dynamics. Notably, we identify a universal regime of Rayleigh–Jeans thermalization where an optical reaction at near-zero optical temperatures can promote the complete and entropically irreversible conversion of light to the fundamental mode at a target frequency. Our theoretical results are corroborated by numerical simulations in settings where second-harmonic generation, sum-frequency generation and four-wave mixing processes can manifest.

Optics

A multimodal large language model for materials science

Understanding and predicting the properties of inorganic materials is crucial for accelerating advancements in materials science and driving applications in energy, electronics and beyond. Integrating material structure data with language-based information through multimodal large language models (LLMs) offers great potential to support these efforts by enhancing human–artificial intelligence interaction. However, a key challenge lies in integrating atomic structures at full resolution into LLMs. In this work, we introduce MatterChat, a versatile structure-aware multimodal LLM that unifies material structural data and textual inputs into a single cohesive model. MatterChat uses a bridging module to effectively align a pretrained universal machine learning interatomic potential with a pretrained LLM, reducing training costs and enhancing flexibility. Our results demonstrate that MatterChat greatly improves performance in material property prediction and human–artificial intelligence interaction, surpassing general-purpose LLMs such as GPT-4. We also demonstrate its usefulness in applications such as more advanced scientific reasoning and step-by-step material synthesis.

Tang, Yingheng [Lawrence Berkeley National Laborat

Multimodal correlative study of Hall transport and magnetic phases in Fe/Gd multilayer systems

The Fe/Gd multilayer system hosts a number of magnetic phases, such as stripe, mixed stripe and skyrmion, skyrmion lattice, and isolated skyrmions for a wide range of temperature and magnetic field. We report different Hall transport signals in a Fe/Gd system through multimodal correlative resonant soft x-ray scattering (RSXS), Hall effect, magneto-optic Kerr effect, and transmission x-ray microscopy measurements. The simultaneous nature of the RSXS and Hall transport measurements allowed us to accurately connect various features in the transport data with the specific magnetic phases. We found that the topological Hall effect (THE) shows peaks with opposite signs, which we attribute to two different mechanisms. Our multimodal correlative study indicates that the sign reversal in THE occurs when the system transforms to and from a skyrmion lattice and low density isolated skyrmion phases. We propose that the skyrmion lattice contributes to the THE through a Berry phase induced emergent magnetic field mechanism in one case, and a skew scattering mechanism corresponding to the isolated low density skyrmion state.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC

Shock tube simulations for the three-layer Richtmyer–Meshkov instability with single-mode and multimode perturbations

While the canonical two-component, single-mode Richtmyer–Meshkov instability (RMI) has been extensively studied, relatively less work has focused on the effects of an additional intermediate-density middle layer. This work investigates such three-material RMI configurations at two Atwood number scenarios using the ARES hydrodynamics code. After validation against previous experimental and computational studies, setups corresponding to recent three-layer shock tube experiments are simulated. Cases with both single-mode and multimode perturbations are studied to quantify mixing across the interface between the materials with highest and intermediate density. In particular, this work is able to comprehensibly examine differences between two- and three-dimensional setups for the single-mode and multimode problems. Observations from previous two-layer investigations still apply in the three-layer setup, but over the time horizons considered, there appears to be insufficient nonlinear mode coupling to create significant differences between two- and three-dimensional simulations following the first passage of a shock. Finally, additional reshock simulations have additional nonlinear growth that does result in expected differences between two- and three-dimensional cases in this three-layer setup, but significant differences do not manifest during the time horizon studied.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC

Rolling Root Mean Square Based Multimodal Anomaly Detection for Real Time Monitoring of Smart Grid

Reliable real-time monitoring is valuable for maintaining the operational integrity of modern electrical smart grids. Deployment of heterogeneous sensing technologies in substations has enabled high-resolution, multichannel waveform monitoring, but also introduces challenges for anomaly detection due to noise, baseline drift, and modality-dependent signal characteristics. In this work, we present a computationally efficient unsupervised method for multimodal event detection based on Rolling Root Mean Square based Event Detection (RRMSED). The method is developed using in-house, field deployed sensors collecting data at a utility substation. The sensing system comprises voltage and current sensors, triaxial accelerometers, and magnetometers, collectively capturing electrical, vibrational, and magnetic waveform measurements at high temporal resolution. RRMSED operates by extracting rolling RMS energy features and their first-order temporal differences from consecutive waveform segments for each channel and then applying channel-specific statistical thresholds learned from historical data. A persistence-based exceedance logic is employed to robustly identify transient events while suppressing impulsive noise, and to provide precise temporal localization with high resolution. The framework is designed for continuous server-side operation and can be deployed in real time without requiring complex models. Experiments on simulated waveform data with known ground truth demonstrate low false positive (FP) and false negative (FN) rates. Application to real substation data shows RRMSED to identify events that are not captured by conventional monitoring indicators including fast transient detection algorithm currently deployed in the system. These results indicate that rolling RMS based features provide an effective and practical basis for real-time multimodal event detection in smart-grid substations.

Mukherjee, Subrata [ORNL] (ORCID:0000000309930338)

Iterative Reconstruction for Multimodal Neutron Tomography

Here, we describe a unified framework for model-based iterative 3-D reconstruction of multimodal neutron transmission, hydrogen-scatter, and induced-fission images from low resolution data recorded using 14.1-MeV neutrons and the associated-particle imaging (API) technique. The framework, which was developed to facilitate use in challenging field-deployment scenarios, is centered around physics-based system models and a total variation (TV) constrained implementation of the simultaneous iterative reconstruction technique (SIRT). Modified to solve a statistically weighted least squares (WLS) problem, the SIRT algorithm is accelerated using ordered subsets and Nesterov’s momentum for which we derive a near-optimal value of the governing Lipschitz constant. The approach enables the reconstruction of images that are high resolution compared to the acquired data and is robust to both limited statistics and a limited number of projection angles. Moreover, the framework is fast enough to be practical. Example images are provided that demonstrate both the ability to perform fast-neutron imaging of high-atomic-number materials with low radiation dose and the benefit of multimodal neutron imaging to identify key materials.

Hydrogen scatter

4D Multimodal Co-attention Fusion Network with Latent Contrastive Alignment for Alzheimer’s Diagnosis

Multimodal neuroimaging provides complementary structural and functional insights into both human brain organization and disease-related dynamics. Recent studies demonstrate enhanced diagnostic sensitivity for Alzheimer’s disease (AD) through synergistic integration of neuroimaging data (e.g., sMRI, fMRI) with tabular data (e.g., behavioral and cognitive tests). However, the intrinsic heterogeneity across modalities (e.g., 4D spatiotemporal fMRI dynamics vs. 3D anatomical sMRI structure) presents critical challenges for discriminative feature fusion, often leading to information loss or biased fusion. To bridge this gap, we propose M2M-AlignNet: a multimodal co-attention network with latent alignment for early AD diagnosis using sMRI and fMRI. At the core of our approach is a multi-patch-to-multi-patch (M2M) contrastive loss function that quantifies and reduces representational discrepancies via weighted patch correspondence, explicitly aligning fMRI components across brain regions with their sMRI structural substrates without one-to-one constraints. Additionally, we propose a latent-as-query co-attention module to autonomously discover fusion patterns, circumventing modality prioritization biases while minimizing feature redundancy. We conduct extensive experiments to confirm the effectiveness of our method and highlight the correspondence between fMRI and sMRI as AD biomarkers.

Wei, Yuxiang [Georgia Institute of Technology]

Emerging Flexible Designs for Geospatial Multimodal Foundation Models

Foundation models are rapidly transforming Earth observation by enabling scalable pretraining across diverse unlabeled geospatial modalities. However, their architectural diversity—ranging from encoder-only to encoder-decoder and masked autoencoding paradigms—makes it challenging to assess performance trade-offs in a consistent manner. In this work, we present an apples-to-apples comparison of leading FM architectures designed for geospatial multimodal reasoning, with a particular focus on flexibility across varied spectral band configurations. We standardize pretraining using identical self-supervised learning objectives and training datasets, and evaluate all models under consistent parameterization on the GEOBench benchmark across classification and segmentation tasks. Our results offer new insights into the design trade-offs between model flexibility, modality alignment, and downstream task performance. By highlighting architectural strengths and limitations under controlled conditions, this study provides practical guidance for building next-generation geospatial foundation models capable of robust multimodal reasoning.

Ambrozio Dias, Philipe [ORNL] (ORCID:0000000194277

Multimodal Atomic Force Microscopy for the Characterization of Metallic Particulates

This study investigates the utility of multimodal, or functional, atomic force microscopy (AFM) for the characterization of individual particles and particle ensembles. In single-particle analyses, AFM imaging modes provided orthogonal insights into mechanical and magnetic properties that are not accessible through conventional electron microscopy. Although the experiments were inherently delicate and time-intensive, with an observed sample loss rate of approximately 30%, these techniques enabled the qualitative differentiation of grain structures and magnetic domains, underscoring the challenges of experimental robustness. For particle ensembles, AFM enabled the extraction of reliable 2D and 3D particle size distributions. However, efforts to chemically differentiate particles based on mechanical property contrasts were limited by scaling effects and the qualitative nature of the data. Overall, multimodal AFM offers valuable complementary information, but its application requires careful consideration of methodological limitations, particularly for sample integrity, calibration, and scalability of mechanical and magnetic measurements. This report is organized to first address the application of AFM to single-particle analysis using various imaging modalities, followed by its potential role in ensemble-level particle screening and differentiation.

36 MATERIALS SCIENCE

Advancing Industry 4.0: Multimodal Sensor Fusion for AI-Based Fault Detection in 3D Printing

Additive manufacturing, particularly fused deposition modeling, is transforming modern production by enabling rapid prototyping and complex part fabrication. However, its layer-by-layer process remains vulnerable to faults such as nozzle clogging, filament runout, and layer misalignment, which compromise print quality and reliability. Traditional inspection methods are costly, time-intensive, and often limited to post-process analysis, making them unsuitable for real-time intervention. In this current study, the authors developed a novel, low-cost, and portable faultdetection system that leverages multimodal sensor fusion and artificial intelligence for real-time monitoring in FDM-based 3D printing. The system integrates acoustic, vibration, and thermal sensing into a non-intrusive architecture, capturing complementary data streams that reflect both mechanical and process-related anomalies. Acoustic and thermal sensors operate in a fully contactless manner, while the vibration sensor requires minimal attachment such that it will not interfere with printer hardware, thereby preserving portability and ease of deployment. The multimodal signals are processed into spectrograms and time-frequency features, which are classified using convolutional neural networks for intelligent fault detection. The proposed system advances Industry 4.0 objectives by offering an affordable, scalable, and practical monitoring solution that improves faultdetection accuracy, reduces waste, and supports sustainable, adaptive manufacturing.

42 ENGINEERING

CAMFeND: Credibility-Aware Multimodal Fake News Detection with Rotational Attention

In the evolving digital landscape, fake news is a significant challenge, influencing public perception and decision-making. Traditional detection approaches focus on single-modal data or simple multimodal fusion, often overlooking deeper interactions and news credibility. We propose a novel model addressing these limitations by introducing rotational attention and news domain information as a feature. Unlike static attention mechanisms, our rotational attention dynamically shifts query, key, and value roles across text and image inputs, enabling richer cross-modal interaction. Incorporating news domain information further enhances the model’s reliability by associating news posts with top domains extracted from Google search results, reducing false detections. This approach assesses both the content and the broader web context in which the news is discussed. Our model outperforms existing state-of-the-art methods by providing deeper, layered multimodal integration and domain information analysis, resulting in a more robust and adaptive fake news detection system.

Gupta, Nidhi

A Dynamic Hierarchical Attention Framework for Multimodal Malware Detection

The increasing use of Android in the worldwide mobile ecosystem has come along with a significant increase in advanced malware, highlighting the critical necessity for efficient, scalable, and adaptable detection systems. Despite recent advancements in machine learning improving malware detection, the majority of current solutions are limited to one, two, or three data modalities, hence neglecting the comprehensive behavioral spectrum of contemporary multi-vector threats. This thesis presents the first comprehensive multimodal framework for Android malware detection, which combines textual, time-series (temporal), graph-based (structural), and visual information using an innovative hierarchical attention mechanism and Dynamic Fusion Controller (DFC). Our methodology consistently classifies and processes modalities as either sequential or structural, facilitating content-adaptive weighting and resilient cross-modal representation learning. We advance the implementation of cutting-edge time series techniques, such as MiniRocket, for malware detection, hence creating new opportunities for temporal analysis in cybersecurity. Comprehensive experimental assessment shows that our framework performs exceptionally well, with 99.46% classification accuracy and 97.15% detection accuracy, significantly outperforming existing approaches through effective multimodal integration and hierarchical attention mechanisms.

Nazmin, Tamanna

Quantifying Operational Drivers of Multimodal Biometric Verification in Aerial Surveillance

Multimodal biometric verification is increasingly applied across operational contexts ranging from close-range security cameras and building-mounted surveillance to long-range ground sensors and unmanned aerial system (UAS) imagery. Variations in acquisition conditions—such as image resolution, viewing geometry, and motion artifacts—pose significant challenges for cross-domain algorithmic generalization. This study evaluates two independent multimodal biometric verification systems developed under the Intelligence Advanced Research Projects Activity (IARPA) Biometric Recognition and Identification at Altitude and Range (BRIAR) program, comparing performance on close-range and aerial datasets. Close-range video served as a baseline to quantify the decline in verification performance on aerial footage. The dataset included six UAS platforms, spanning small quadcopters at 10m altitude to medium-sized fixed-wing aircraft at 360m. Mixed-effects logistic regression identified image resolution (head and body pixel counts), head height, sensor characteristics, and algorithm selection as primary determinants of verification success, whereas demographic attributes and mission gait were not significant predictors. Activity type and collection site influenced performance in close-range data but had negligible impact on UAS imagery. These results clarify modality-specific strengths and limitations and highlight opportunities to enhance cross-domain biometric verification.

Peluso, Alina [ORNL] (ORCID:0000000328950406)