Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “hierarchical sampling”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

Hierarchical Modeling to Enhance Spectrophotometry Measurements—Overcoming Dynamic Range Limitations for Remote Monitoring of Neptunium

A robust hierarchical model has been demonstrated for monitoring a wide range of neptunium concentrations (0.75–890 mM) and varying temperatures (10–80 °C) using chemometrics and feature selection. The visible–near infrared electronic absorption spectrum (400–1700 nm) of monocharged neptunyl dioxocation (Np(V) = NpO2+) includes many bands, which have molar absorption coefficients that differ by nearly 2 orders of magnitude. The shape, position, and intensity of these bands differ with chemical interactions and changing temperature. These challenges make traditional quantification by univariate methods unfeasible. Measuring Np(V) concentration over several orders of magnitude would typically necessitate cells with varying path length, optical switches, and/or multiple spectrophotometers. Alternatively, the differences in the molar extinction coefficients for multiple absorption bands can be used to quantify Np(V) concentration over 3 orders of magnitude with a single optical path length (1 mm) and a hierarchical multivariate model. In this work, principal component analysis was used to distinguish the concentration regime of the sample, directing it to the relevant partial least squares regression submodels. Each submodel was optimized with unique feature selection filters that were selected by a genetic algorithm to enhance predictions. Through this approach, the percent root mean square error of prediction values were ≤1.05% for Np(V) concentrations and ≤4% for temperatures. This approach may be applied to other nuclear fuel cycle and environmental applications requiring real-time spectroscopic measurements over a wide range of conditions.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Formation Flying Control Implementation in Highly Elliptical Orbits

The Tschauner-Hempel equations are widely used to correct the separation distance drifts between a pair of satellites within a constellation in highly elliptical orbits [1]. This set of equations was discretized in the true anomaly angle [1] to be used in a digital steady-state hierarchical controller [2]. This controller [2] performed the drift correction between a pair of satellites within the constellation. The objective of a discretized system is to develop a simple algorithm to be implemented in the computer onboard the satellite. The main advantage of the discrete systems is that the computational time can be reduced by selecting a suitable sampling interval. For this digital system, the amount of data will depend on the sampling interval in the true anomaly angle [3]. The purpose of this paper is to implement the discrete Tschauner-Hempel equations and the steady-state hierarchical controller in the computer onboard the satellite. This set of equations is expressed in the true anomaly angle in which a relation will be formulated between the time and the true anomaly angle domains.

Capo-Lugo, Pedro A.↗

Inorganic characterization of switchgrass biomass using laser-induced breakdown spectroscopy

The inorganic characterization of 74 samples of switchgrass using laser-induced breakdown spectroscopy (LIBS) was undertaken. Determination of ash and inorganic elements content in biomass materials is vital for feedstock screening for bioconversion processes. Hierarchical models using principal component analysis (PCA) and partial least square analysis (PLS) were used to determine the presence of specific elemental micronutrients that are important in determining plant health for robust biomass production. LIBS uses a 532 nm laser with 45 mJ of laser power to excite the samples of switchgrass plant material and the emission of all the elements present in the plant samples were recorded in single spectra with a wide wavelength range of 200–800 nm. The results were compared to the laboratory standard technique, e.g., ICP-OES technique, to determine the true values for major micronutrients such as, silicon (Si), potassium (K), calcium (Ca), magnesium (Mg), phosphorus (P), and sulfur (S). Overall, our objectives were: 1) To determine the spectral features of switchgrass containing different amounts of these elements and 2) To examine the viability of this technique for determining the quality of the feedstock in terms of its inorganic composition. Cross-validation results showed that the broad-based model developed is promising for inorganics prediction in switchgrass. The LIBS validation prediction for the micronutrient elements mentioned here have been obtained. The regression coefficients for Si, were obtained to be 0.995, 0.994 for calibration and validation respectively, in case of Ca the regression coefficients were, 0.994 and 0.992 for calibration and validation. Similarly, in the case of Mg and K these were calculated to be 0.992 and 0.985, and 0.994 and 0.993 respectively. The regression coefficients are not as good as those for the elements mentioned, in case of the two elements S and P. They are 0.957, and 0.878, and 0.952 and 0.894 respectively for calibration, validation for the two elements. This demonstrates that LIBS-based techniques are inherently well suited for diverse environmental applications. Furthermore, LIBS along with PLS model can show capability in determining the viability of switchgrass as a biomass in the production of biofuels and survivability of switchgrass in processes associated with climate change. LIBS can help determining which switchgrass would be appropriate for a specific conversion process that favors low ash content overall or low value of specific inorganics.

59 BASIC BIOLOGICAL SCIENCES↗

Stochastic modeling and statistical calibration with model error and scarce data

This paper introduces a procedure to assess the predictive accuracy of stochastic models subject to model error and sparse data. Model error is introduced as uncertainty on the coefficients of appropriate polynomial chaos expansions (PCE). The error associated with finite sample size allows us to conceive of these coefficients as statistics of the data that we describe as random variables whose influence on output quantities of interest is evaluated through the extended polynomial chaos expansion (EPCE). A Bayesian data assimilation scheme is introduced to update these expansions by considering the resulting nested chaos expansion as a hierarchical probabilistic model. Stochastic models of quantities of interest (QoI) are thus constructed and efficiently evaluated. Here, the Metropolis–Hastings Markov chain Monte Carlo procedure is used to sample the posterior. Two illustrative analytical and numerical problems are used to demonstrate the proposed approach.

Bayesian inference↗

Multiscale Flow for robust and optimal cosmological analysis

We propose Multiscale Flow, a generative Normalizing Flow that creates samples and models the field-level likelihood of two-dimensional cosmological data such as weak lensing. Multiscale Flow uses hierarchical decomposition of cosmological fields via a wavelet basis and then models different wavelet components separately as Normalizing Flows. The log-likelihood of the original cosmological field can be recovered by summing over the log-likelihood of each wavelet term. This decomposition allows us to separate the information from different scales and identify distribution shifts in the data such as unknown scale-dependent systematics. The resulting likelihood analysis can not only identify these types of systematics, but can also be made optimal, in the sense that the Multiscale Flow can learn the full likelihood at the field without any dimensionality reduction. We apply Multiscale Flow to weak lensing mock datasets for cosmological inference and show that it significantly outperforms traditional summary statistics such as power spectrum and peak counts, as well as machine learning–based summary statistics such as scattering transform and convolutional neural networks. We further show that Multiscale Flow is able to identify distribution shifts not in the training data such as baryonic effects. Finally, we demonstrate that Multiscale Flow can be used to generate realistic samples of weak lensing data.

79 ASTRONOMY AND ASTROPHYSICS↗

Alteration mapping at Goldfield, Nevada, by cluster and discriminant analysis of LANDSAT digital data

The ability of Landsat multispectral digital data to differentiate among 62 combinations of rock and alteration types at the Goldfield mining district of Western Nevada was investigated by using statistical techniques of cluster and discriminant analysis. Multivariate discriminant analysis was not effective in classifying each of the 62 groups, with classification results essentially the same whether data of four channels alone or combined with six ratios of channels were used. Bivariate plots of group means revealed a cluster of three groups including mill tailings, basalt and all other rock and alteration types. Automatic hierarchical clustering based on the fourth dimensional Mahalanobis distance between group means of 30 groups having five or more samples was performed. The results of the cluster analysis revealed hierarchies of mill tailings vs. natural materials, basalt vs. non-basalt, highly reflectant rocks vs. other rocks and exclusively unaltered rocks vs. predominantly altered rocks. The hierarchies were used to determine the order in which sets of multiple discriminant analyses were to be performed and the resulting discriminant functions were used to produce a map of geology and alteration which has an overall accuracy of 70 percent for discriminating exclusively altered rocks from predominantly altered rocks.

Ballew, G.↗

Alteration mapping at Goldfield, Nevada, by cluster and discriminant analysis of Landsat digital data

The ability of Landsat multispectral digital data to differentiate among 62 combinations of rock and alteration types at the Goldfield mining district of Western Nevada was investigated by using statistical techniques of cluster and discriminant analysis. Multivariate discriminant analysis was not effective in classifying each of the 62 groups, with classification results essentially the same whether data of four channels alone or combined with six ratios of channels were used. Bivariate plots of group means revealed a cluster of three groups including mill tailings, basalt and all other rock and alteration types. Automatic hierarchical clustering based on the fourth dimensional Mahalanobis distance between group means of 30 groups having five or more samples was performed using Johnson's HICLUS program. The results of the cluster analysis revealed hierarchies of mill tailings vs. natural materials, basalt vs. non-basalt, highly reflectant rocks vs. other rocks and exclusively unaltered rocks vs. predominantly altered rocks. The hierarchies were used to determine the order in which sets of multiple discriminant analyses were to be performed and the resulting discriminant functions were used to produce a map of geology and alteration which has an overall accuracy of 70 percent for discriminating exclusively altered rocks from predominantly altered rocks.

Ballew, G.↗

Learning In networks

Intelligent systems require software incorporating probabilistic reasoning, and often times learning. Networks provide a framework and methodology for creating this kind of software. This paper introduces network models based on chain graphs with deterministic nodes. Chain graphs are defined as a hierarchical combination of Bayesian and Markov networks. To model learning, plates on chain graphs are introduced to model independent samples. The paper concludes by discussing various operations that can be performed on chain graphs with plates as a simplification process or to generate learning algorithms.

Buntine, Wray L.↗

Probabilisitc Geobiological Classification Using Elemental Abundance Distributions and Lossless Image Compression in Recent and Modern Organisms

Last year we presented techniques for the detection of fossils during robotic missions to Mars using both structural and chemical signatures[Storrie-Lombardi and Hoover, 2004]. Analyses included lossless compression of photographic images to estimate the relative complexity of a putative fossil compared to the rock matrix [Corsetti and Storrie-Lombardi, 2003] and elemental abundance distributions to provide mineralogical classification of the rock matrix [Storrie-Lombardi and Fisk, 2004]. We presented a classification strategy employing two exploratory classification algorithms (Principal Component Analysis and Hierarchical Cluster Analysis) and non-linear stochastic neural network to produce a Bayesian estimate of classification accuracy. We now present an extension of our previous experiments exploring putative fossil forms morphologically resembling cyanobacteria discovered in the Orgueil meteorite. Elemental abundances (C6, N7, O8, Na11, Mg12, Ai13, Si14, P15, S16, Cl17, K19, Ca20, Fe26) obtained for both extant cyanobacteria and fossil trilobites produce signatures readily distinguishing them from meteorite targets. When compared to elemental abundance signatures for extant cyanobacteria Orgueil structures exhibit decreased abundances for C6, N7, Na11, All3, P15, Cl17, K19, Ca20 and increases in Mg12, S16, Fe26. Diatoms and silicified portions of cyanobacterial sheaths exhibiting high levels of silicon and correspondingly low levels of carbon cluster more closely with terrestrial fossils than with extant cyanobacteria. Compression indices verify that variations in random and redundant textural patterns between perceived forms and the background matrix contribute significantly to morphological visual identification. The results provide a quantitative probabilistic methodology for discriminating putatitive fossils from the surrounding rock matrix and &om extant organisms using both structural and chemical information. The techniques described appear applicable to the geobiological analysis of meteoritic samples or in situ exploration of the Mars regolith. Keywords: cyanobacteria, microfossils, Mars, elemental abundances, complexity analysis, multifactor analysis, principal component analysis, hierarchical cluster analysis, artificial neural networks, paleo-biosignatures

Storrie-Lombardi, Michael C.↗

Advancing Chemical Separations: Unraveling the Structure and Dynamics of Phase Splitting in Liquid–Liquid Extraction

Liquid-liquid extraction (LLE), the go-to process for a variety of chemical separations, is limited by spontaneous organic phase splitting upon sufficient solute loading, called third phase formation. In this study we explore the applicability of critical phenomena theory to gain insight into this deleterious phase behavior with the goal of improving separations efficiency and minimizing waste. Here, a series of samples representative of rare earth purification were constructed to include each of one light and one heavy lanthanide (cerium and lutetium) paired with one of two common malonamide extractants (DMDOHEMA and DMDBTDMA). The resulting postextraction organic phases are chemically complex and often form rich hierarchical structures whose statics and dynamics near the critical point were probed herein with small-angle X-ray scattering and high-speed X-ray photon correlation spectroscopy. Despite their different extraction behaviors, all samples show remarkably similar critical behavior with exponents well described by classical critical point theory consistent with the 3D Ising model, where the critical behavior is characterized by fluctuations with a single diverging length scale. This unexpected result indicates a significant reduction in relevant chemical parameters at the critical point, indicating that the underlying behavior of phase transitions in LLE rely on far fewer variables than are generally assumed. The obtained scalar order parameter is attributed to the extractant fraction of the extractant/diluent mixture, revealing that other solution components and their respective concentrations simply shift the critical temperature but do not affect the nature of the critical fluctuations. These findings point to an opportunity to drastically simplify studies of liquid-liquid phase separation and phase diagram development in general while providing insights into LLE process improvement.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Do the major axes of rich clusters of galaxies point toward their neighbors?

The major axis orientation of rich clusters of galaxies, determined from an analysis of X-ray images, is used to investigate whether these clusters point toward their nearest neighbors. No statistical significance is found for a pointing effect between clusters and their nearest neighbors in either X-ray, optical, or combined X-ray and optical samples. Using updated redshifts and permitting nonstatistical sample Abell clusters as nearest neighbors does not affect this conclusion. The lack of statistical significance for a pointing effect favors hierarchical models in which galaxies form first, followed by clusters and then superclusters. For clusters with well-defined X-ray orientations, it is found that cluster position angles determined from X-ray and optical data are in general agreement.

Ulmer, M. P.↗

A petabyte size electronic library using the N-Gram memory engine

A model library containing petabytes of data is proposed by Triada, Ltd., Ann Arbor, Michigan. The library uses the newly patented N-Gram Memory Engine (Neurex), for storage, compression, and retrieval. Neurex splits data into two parts: a hierarchical network of associative memories that store 'information' from data and a permutation operator that preserves sequence. Neurex is expected to offer four advantages in mass storage systems. Neurex representations are dense, fully reversible, hence less expensive to store. Neurex becomes exponentially more stable with increasing data flow; thus its contents and the inverting algorithm may be mass produced for low cost distribution. Only a small permutation operator would be recalled from the library to recover data. Neurex may be enhanced to recall patterns using a partial pattern. Neurex nodes are measures of their pattern. Researchers might use nodes in statistical models to avoid costly sorting and counting procedures. Neurex subsumes a theory of learning and memory that the author believes extends information theory. Its first axiom is a symmetry principle: learning creates memory and memory evidences learning. The theory treats an information store that evolves from a null state to stationarity. A Neurex extracts information data without a priori knowledge; i.e., unlike neural networks, neither feedback nor training is required. The model consists of an energetically conservative field of uniformly distributed events with variable spatial and temporal scale, and an observer walking randomly through this field. A bank of band limited transducers (an 'eye'), each transducer in a bank being tuned to a sub-band, outputs signals upon registering events. Output signals are 'observed' by another transducer bank (a mid-brain), except the band limit of the second bank is narrower than the band limit of the first bank. The banks are arrayed as n 'levels' or 'time domains, td.' The banks are the hierarchical network (a cortex) and transducers are (associative) memories. A model Neurex was built and studied. Data were 50 MB to 10 GB samples of text, data base, and images: black/white, grey scale, and high resolution in several spectral bands. Memories at td, S(m(sub td)), were plotted against outputs of memories at td-1. S(m(sub td)) was Boltzman distributed, and memory frequencies exhibited self-organized criticality (SOC); i.e., 'l/f(sup beta)' after long exposures to data. Whereas output signals from level n may be encoded with B(sub output) = O(-log(2)f(sup beta)) bits, and input data encoded with B(sub input) = O((S(td)/S(td-1))(sup n)), B(sup output)/B(sub input) is much less than 1 always, the Neurex determines a canonical code for data and it is a lossless data compressor. Further tests are underway to confirm these results with more data types and larger samples.

Bugajski, Joseph M.↗

Probabilistic Classification Using Elemental Abundance Distributions and Lossless Image Compression in Apollo 17 Lunar Dust Samples from Mare Serenitatis

We have previously outlined a strategy for the detection of fossils [Storrie-Lombardi and Hoover, 2004] and extant microbial life [Storrie-Lombaudi and Hoover, 20051 during robotic missions to Mars using co-registered structural and chemical signatures. Data inputs included image lossless compression indices to estimate relative textural complexity and elemental abundance distributions. Two exploratory classification algorithms (principal component analysis and hierarchical cluster analysis) provide an initial tentative classification of all targets. Nonlinear stochastic neural networks are then trained to produce a Bayesian estimate of algorithm classification accuracy. The strategy previously has been successful in distinguishing regions of biotic and abiotic alteration of basalt glass from unaltered samples. [Storrie-Lombardi and Fisk, 2004; Storrie-Lombardi and Fisk, 2004] Such investigations of abiotic versus biotic alteration of terrestrial mineralogy on Earth are compromised by .the difficulty finding mineralogy completely unaffected by the ubiquitous presence of microbial life on the planet. The renewed interest in lunar exploration offers an opportunity to investigate geological materials that may exhibit signs of aqueous alteration, but are highly unlikely to contain contaminating biological weathering signatures. We here present an extension of our earlier data set to include lunar dust samples obtained during the Apollo 17 mission. Apollo 17 landed in the Taurus-Littrow Valley in Mare Serenitatis. Most of the rock samples from this region of the lunar highlands are basalts comprised primarily of plagioclase and pyroxene and selected examples of orange and black volcanic glass. SEM images and elemental abundances (C6, N7, O8, Na11, Mg12, Al13, Si14, P15, S16, Cll7, K19, Ca20, Fe26) for a series of targets in the lunar dust samples are compared to the extant cyanobacteria, fossil trilobites, Orgueil meteorite, and terrestrial basalt targets previously discussed. The data set provides a first step in producing a quantitative probabilistic methodology for geobiological analysis of returned lunar samples or in situ exploration.

Storrie-Lombardi, Michael C.↗

Untangling the Galaxy. II. Structure within 3 kpc

We present the results of the hierarchical clustering analysis of the Gaia DR2 data to search for clusters, comoving groups, and other stellar structures. The current paper builds on the sample from the previous work, extending it in distance from 1 to 3 kpc and increasing the number of identified structures up to 8292. To aid in the analysis of the population properties, we developed a neural network called Auriga to robustly estimate the age, extinction, and distance of a stellar group based on the input photometry and parallaxes of the individual members. We apply Auriga to derive the properties of not only the structures found in this paper, but also previously identified open clusters. Through this work, we examine the temporal structure of the spiral arms. Specifically, we find that the Sagittarius Arm has moved by >500 pc in the last 100 Myr and the Perseus Arm has been experiencing a relative lull in star formation activity over the last 25 Myr. We confirm the findings of the previous paper on the transient nature of the spiral arms, with the timescale of transition of a few 100 Myr. Finally, we find a peculiar ~1 Gyr old stream of stars that appears to be heliocentric. Its origin is unclear.

79 ASTRONOMY AND ASTROPHYSICS↗

Ocean Data from MODIS at the NASA Goddard DAAC

Terra satellite carrying the Moderate Resolution Imaging Spectroradiometer (MODIS) was successfully launched on December 18, 1999. Some of the 36 different wavelengths that MODIS samples have never before been measured from space. New ocean data products, which have not been derived on a global scale before, are made available for research to the scientific community. For example, MODIS uses a new split window in the four-micron region for the better measurement of Sea Surface Temperature (SST), and provides the unprecedented ability (683 nm band) to measure chlorophyll fluorescence. At full ocean production, more than a thousand different ocean products in three major categories (ocean color, sea surface temperature, and ocean primary production) are archived at the NASA Goddard Earth Sciences (GES) Distributed Active Archive Center (DAAC) at the rate of approx. 230GB/day. The challenge is to distribute such large volumes of data to the ocean community. It is achieved through a combination of public and restricted EOS Data Gateways, the GES DAAC Search and Order WWW interface, and an FTP site that contains samples of MODIS data. A new Search and Order WWW interface at http://acdisx.gsfc.nasa.gov/data/ developed at the GES DAAC is based on a hierarchical organization of data, will always return non-zero results. It has a very convenient geographical representation of five-minute data granule coverage for each day MODIS Data Support Team (MDST) continues the tradition of quality support at the GES DAAC for the ocean color data from the Coastal Zone Color Scanner (CZCS) and the Sea Viewing Wide Field-of-View Sensor (SeaWiFS) by providing expert assistance to users in accessing data products, information on visualization tools, documentation for data products and formats (Hierarchical Data Format-Earth Observing System (HDF-EOS)), information on the scientific content of products and metadata. Visit the MDST website at http://daac.gsfc.nasa.gov/CAMPAIGN DOCS/MODIS/index.html

Leptoukh, Gregory G.↗

Remote Sensing of the Urban Heat Island Effect Across Biomes in the Continental USA

Impervious surface area (ISA) from the Landsat TM-based NLCD 2001 dataset and land surface temperature (LST) from MODIS averaged over three annual cycles (2003-2005) are used in a spatial analysis to assess the urban heat island (UHI) skin temperature amplitude and its relationship to development intensity, size, and ecological setting for 38 of the most populous cities in the continental United States. Development intensity zones based on %ISA are defined for each urban area emanating outward from the urban core to the nonurban rural areas nearby and used to stratify sampling for land surface temperatures and NDVI. Sampling is further constrained by biome and elevation to insure objective intercomparisons between zones and between cities in different biomes permitting the definition of hierarchically ordered zones that are consistent across urban areas in different ecological setting and across scales. We find that ecological context significantly influences the amplitude of summer daytime UHI (urban-rural temperature difference) the largest (8 C average) observed for cities built in biomes dominated by temperate broadleaf and mixed forest. For all cities combined, ISA is the primary driver for increase in temperature explaining 70% of the total variance in LST. On a yearly average, urban areas are substantially warmer than the non-urban fringe by 2.9 C, except for urban areas in biomes with arid and semiarid climates. The average amplitude of the UHI is remarkably asymmetric with a 4.3 C temperature difference in summer and only 1.3 C in winter. In desert environments, the LST's response to ISA presents an uncharacteristic "U-shaped" horizontal gradient decreasing from the urban core to the outskirts of the city and then increasing again in the suburban to the rural zones. UHI's calculated for these cities point to a possible heat sink effect. These observational results show that the urban heat island amplitude both increases with city size and is seasonally asymmetric for a large number of cities across most biomes. The implications are that for urban areas developed within forested ecosystems the summertime UHI can be quite high relative to the wintertime UHI suggesting that the residential energy consumption required for summer cooling is likely to increase with urban growth within those biomes.

Imhoff, Marc L.↗

Tailoring plasticity mechanisms in compositionally graded hierarchical steels fabricated using additive manufacturing

Abstract While there exists in nature abundant examples of materials with site-specific gradients in microstructures and properties, engineers and designers have traditionally used monolithic materials with discrete properties. Now, however, additive manufacturing (AM) offers the possibility of creating structures that mimic some aspects of nature. One example that has attracted attention in the recent years is the hierarchical structure in bamboo. The hierarchical architecture in bamboo is characterized by spatial gradients in properties and microstructures and is well suited to accommodate and survive complex stress states, severe mechanical forces, and large deformations. While AM has been used routinely to fabricate functionally graded materials, this study distinguishes itself by leveraging AM and physical metallurgy concepts to trigger cascading deformation in a single sample. Specifically, we have been successful in using AM to fabricate steel with unique spatial hierarchies in structure and property to emulate the structure and deformation mechanisms in natural materials. This study shows an improvement in the strength and ductility of the nature-inspired “hierarchical steel” compared with conventional cast stainless steels. In situ characterization proves that this improvement is due to the sequential activation of multiple deformation mechanisms namely twinning, transformation-induced plasticity, and dislocation-based plasticity. While significantly higher strengths can be achieved by refining the chemical and processing technique, this study sets the stage to achieve the paradigm of using AM to fabricate structures which emulate the flexibility in mechanical properties of natural materials and are able to adapt to in-service conditions.

36 MATERIALS SCIENCE↗

Hierarchical Multi-agent Large Language Model Reasoning for Autonomous Heterogeneous Catalyst Discovery

Artificial intelligence is reshaping scientific exploration, but most methods automate procedural tasks without engaging in scientific reasoning, limiting autonomy in discovery. We demonstrate that hierarchical agentic large language model reasoning can efficiently drive simulation and scientific exploration. Across two chemical applications, CO adsorption on Cu surface transition metal adatoms and on M–N–C catalysts, reasoning-guided exploration reduces required atomistic simulations by up to 90% relative to heuristic or random selection. Comparisons across single-agent, multi-agent, and stochastic baselines show that hierarchical strategies yield more coherent and information-efficient search trajectories. Reasoning traces reveal chemically grounded decisions that cannot be explained by semantic bias or stochastic sampling. We realize these agentic reasoning strategies in Materials Agents for Simulation and Theory in Electronic-structure Reasoning (MASTER), a multimodal system that translates natural language into density functional theory workflows. Altogether, multi-agent collaboration accelerates heterogeneous catalyst discovery and marks a step toward more autonomous, reasoning-guided scientific exploration.

30 DIRECT ENERGY CONVERSION↗