Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “model discovery”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 127 records · Page 7

A framework for strategic discovery of credible neural network surrogate models under uncertainty

The widespread integration of deep neural networks in developing data-driven surrogate models for high-fidelity simulations of complex physical systems highlights the critical necessity for robust uncertainty quantification techniques and credibility assessment methodologies, ensuring the reliable deployment of surrogate models in consequential decision-making. Here, this study presents the Occam Plausibility Algorithm for surrogate models (OPAL-surrogate), providing a systematic framework to uncover predictive neural network-based surrogate models within the large space of potential models, including various neural network classes and choices of architecture and hyperparameters. The framework is grounded in hierarchical Bayesian inferences and employs model validation tests to evaluate the credibility and prediction reliability of the surrogate models under uncertainty. Leveraging these principles, OPAL-surrogate introduces a systematic and efficient strategy for balancing the trade-off between model complexity, accuracy, and prediction uncertainty. The effectiveness of OPAL-surrogate is demonstrated through two modeling problems, including the deformation of porous materials for building insulation and turbulent combustion flow for ablation of solid fuels within hybrid rocket motors.

42 ENGINEERING↗

DLHub: Simplifying publication, discovery, and use of machine learning models in science

Machine Learning (ML) has become a critical tool enabling new methods of analysis and driving deeper understanding of phenomena across scientific disciplines. There is a growing need for "learning systems" to support various phases in the ML lifecycle. While others have focused on supporting model development, training, and inference, few have focused on the unique challenges inherent in science, such as the need to publish and share models and to serve them on a range of available computing resources. In this paper, we present the Data and Learning Hub for science (DLHub), a learning system designed to support these use cases. Specifically, DLHub enables publication of models, with descriptive metadata, persistent identifiers, and flexible access control. It packages arbitrary models into portable servable containers, and enables low-latency, distributed serving of these models on heterogeneous compute resources. In this work, we show that DLHub supports low-latency model inference comparable to other model serving systems including TensorFlow Serving, SageMaker, and Clipper, and improved performance, by up to 95%, with batching and memoization enabled. We also show that DLHub can scale to concurrently serve models on 500 containers. Finally, we describe five case studies that highlight the use of DLHub for scientific applications.

97 MATHEMATICS AND COMPUTING↗

When ancient numerical demons meet physics-informed machine learning: adjoint-based gradients for implicit differentiable modeling

Recent advances in differentiable modeling, a genre of physics-informed machine learning that trains neural networks (NNs) together with process-based equations, have shown promise in enhancing hydrological models' accuracy, interpretability, and knowledge-discovery potential. Current differentiable models are efficient for NN-based parameter regionalization, but the simple explicit numerical schemes paired with sequential calculations (operator splitting) can incur numerical errors whose impacts on models' representation power and learned parameters are not clear. Implicit schemes, however, cannot rely on automatic differentiation to calculate gradients due to potential issues of gradient vanishing and memory demand. Here we propose a “discretize-then-optimize” adjoint method to enable differentiable implicit numerical schemes for the first time for large-scale hydrological modeling. The adjoint model demonstrates comprehensively improved performance, with Kling–Gupta efficiency coefficients, peak-flow and low-flow metrics, and evapotranspiration that moderately surpass the already-competitive explicit model. Therefore, the previous sequential-calculation approach had a detrimental impact on the model's ability to represent hydrological dynamics. Furthermore, with a structural update that describes capillary rise, the adjoint model can better describe baseflow in arid regions and also produce low flows that outperform even pure machine learning methods such as long short-term memory networks. The adjoint model rectified some parameter distortions but did not alter spatial parameter distributions, demonstrating the robustness of regionalized parameterization. Despite higher computational expenses and modest improvements, the adjoint model's success removes the barrier for complex implicit schemes to enrich differentiable modeling in hydrology.

58 GEOSCIENCES↗

Machine Learning Design of Perovskite Catalytic Properties

Abstract Discovering new materials that efficiently catalyze the oxygen reduction and evolution reactions is critical for facilitating the widespread adoption of solid oxide fuel cell and electrolyzer (SOFC/SOEC) technologies. Here, machine learning (ML) models are developed to predict perovskite catalytic properties critical for SOFC/SOEC applications, including oxygen surface exchange, oxygen diffusivity, and area specific resistance (ASR). The models are based on trivial‐to‐calculate elemental features and are more accurate and dramatically faster than the best models based on ab initio‐derived features, potentially eliminating the need for ab initio calculations in descriptor‐based screening. The model of ASR enables temperature‐dependent predictions, has well calibrated uncertainty estimates and online accessibility. Use of temporal cross‐validation reveals the model to be effective at discovering new promising materials prior to their initial discovery, demonstrating the model can make meaningful predictions. Using the SHapley Additive ExPlanations (SHAP) approach, detailed discussion of different approaches of model featurization is provided for ML property prediction. Finally, the model is used to screen more than 19 million perovskites to develop a list of promising cheap, earth‐abundant, stable, and high performing materials, and find some top materials contain mixtures of less‐explored elements (e.g., K, Bi, Y, Ni, Cu) worth exploring in more detail.

25 ENERGY STORAGE↗

Toward a Holistic Performance Evaluation of Large Language Models Across Diverse AI Accelerators

Artificial intelligence (AI) methods have become critical in scientific applications to help accelerate scientific discovery. Large language models (LLMs) are being considered a promising approach to address some challenging problems because of their superior generalization capabilities across domains. The effectiveness of the models and the accuracy of the applications are contingent upon their efficient execution on the underlying hardware infrastructure. Specialized Al accelerator hardware systems have recently become available for accelerating Al applications. However, the comparative performance of these AI accelerators on large language models has not been previously studied. In this paper, we systematically study LLMs on multiple AI accelerators and GPUs and evaluate their performance characteristics for these models. We evaluate these systems with (i) a micro-benchmark using a core transformer block, (ii) a GPT-2 model, and (iii) an 1,I,M-driven science use case, GenSLM. We present our findings and analyses of the models' performance to better understand the intrinsic capabilities of AI accelerators. Furthermore, our analysis takes into account key factors such as sequence lengths, scaling behavior, and sensitivity to gradient accumulation steps.

Emani, Murali↗

A physics-informed operator regression framework for extracting data-driven continuum models

The application of deep learning toward discovery of data-driven models requires careful application of inductive biases to obtain a description of physics which is both accurate and robust. We present here a framework for discovering continuum models from high fidelity molecular simulation data. Our approach applies a neural network parameterization of governing physics in modal space, allowing a characterization of differential operators while providing structure which may be used to impose biases related to symmetry, isotropy, and conservation form. Here, we demonstrate the effectiveness of our framework for a variety of physics, including local and nonlocal diffusion processes and single and multiphase flows. For the flow physics we demonstrate this approach leads to a learned operator that generalizes to system characteristics not included in the training sets, such as variable particle sizes, densities, and concentration.

42 ENGINEERING↗

Enhancing molecular design efficiency: Uniting language models and generative networks with genetic algorithms

This study examines the effectiveness of generative models in drug discovery, material science, and polymer science, aiming to overcome constraints associated with traditional inverse design methods relying on heuristic rules. Generative models generate synthetic data resembling real data, enabling deep learning model training without extensive labeled datasets. They prove valuable in creating virtual libraries of molecules for material science and facilitating drug discovery by generating molecules with specific properties. While generative adversarial networks (GANs) are explored for these purposes, mode collapse restricts their efficacy, limiting novel structure variability. To address this, we introduce a masked language model (LM) inspired by natural language processing. Although LMs alone can have inherent limitations, we propose a hybrid architecture combining LMs and GANs to efficiently generate new molecules, demonstrating superior performance over standalone masked LMs, particularly for smaller population sizes. This hybrid LM-GAN architecture enhances efficiency in optimizing properties and generating novel samples.

97 MATHEMATICS AND COMPUTING↗

Explaining and predicting human behavior and social dynamics in simulated virtual worlds: reproducibility, generalizability, and robustness of causal discovery methods

Ground Truth program was designed to evaluate social science modeling approaches using simulation test beds with ground truth intentionally and systematically embedded to understand and model complex Human Domain systems and their dynamics Lazer et al. (Science 369:1060–1062, 2020). Our multidisciplinary team of data scientists, statisticians, experts in Artificial Intelligence (AI) and visual analytics had a unique role on the program to investigate accuracy, reproducibility, generalizability, and robustness of the state-of-the-art (SOTA) causal structure learning approaches applied to fully observed and sampled simulated data across virtual worlds. In addition, we analyzed the feasibility of using machine learning models to predict future social behavior with and without causal knowledge explicitly embedded. In this paper, we first present our causal modeling approach to discover the causal structure of four virtual worlds produced by the simulation teams—Urban Life, Financial Governance, Disaster and Geopolitical Conflict. Our approach adapts the state-of-the-art causal discovery (including ensemble models), machine learning, data analytics, and visualization techniques to allow a human-machine team to reverse-engineer the true causal relations from sampled and fully observed data. We next present our reproducibility analysis of two research methods team’s performance using a range of causal discovery models applied to both sampled and fully observed data, and analyze their effectiveness and limitations. We further investigate the generalizability and robustness to sampling of the SOTA causal discovery approaches on additional simulated datasets with known ground truth. Our results reveal the limitations of existing causal modeling approaches when applied to large-scale, noisy, high-dimensional data with unobserved variables and unknown relationships between them. We show that the SOTA causal models explored in our experiments are not designed to take advantage from vasts amounts of data and have difficulty recovering ground truth when latent confounders are present; they do not generalize well across simulation scenarios and are not robust to sampling; they are vulnerable to data and modeling assumptions, and therefore, the results are hard to reproduce. Finally, when we outline lessons learned and provide recommendations to improve models for causal discovery and prediction of human social behavior from observational data, we highlight the importance of learning data to knowledge representations or transformations to improve causal discovery and describe the benefit of causal feature selection for predictive and prescriptive modeling.

97 MATHEMATICS AND COMPUTING↗

Editorial: Predicting near-earth space environment: new perspective and capabilities in the AI age

Editorial on the Research Topic Predicting near-earth space environment: new perspective and capabilities in the AI age The near-Earth space environment is not only an operational hazard for space missions, but also a scientific laboratory for advancing our understanding and prediction of space plasma populations. This Research Topic is organized around three interconnected themes: observational datasets, machine-learning (ML) model development, and the discovery of new physical insights through those models. Its primary goal is to highlight the emerging capabilities in space environment prediction that are enabled, or will be enabled, by integrating advanced techniques—including AI/ML methods—with long-term curated datasets.

58 GEOSCIENCES↗

Weak Form Scientific Machine Learning: Test Function Construction for System Identification

Weak form Scientific Machine Learning (WSciML) is a recently developed framework for data-driven modeling and scientific discovery. It leverages the weak form of equation error residuals to provide enhanced noise robustness in system identification via convolving model equations with test functions, reformulating the problem to avoid direct differentiation of data. The performance, however, relies on wisely choosing a set of compactly supported test functions. In this work, we mathematically motivate a novel data-driven method for constructing Single-scale-Local reference functions for creating the set of test functions. Our approach numerically approximates the integration error introduced by the quadrature and identifies the support size for which the error is minimal, without requiring access to the model parameter values. Through numerical experiments across various models, noise levels, and temporal resolutions, we demonstrate that the selected supports consistently align with regions of minimal parameter estimation error. We also compare the proposed method against the strategy for constructing Multi-scale-Global (and orthogonal) test functions introduced in our prior work, demonstrating the improved computational efficiency.

FOS: Computer and information sciences↗

CASTELO: clustered atom subtypes aided lead optimization—a combined machine learning and molecular modeling method

Background: Drug discovery is a multi-stage process that comprises two costly major steps: pre-clinical research and clinical trials. Among its stages, lead optimization easily consumes more than half of the pre-clinical budget. We propose a combined machine learning and molecular modeling approach that partially automates lead optimization workflow in silico, providing suggestions for modification hot spots. Results: The initial data collection is achieved with physics-based molecular dynamics simulation. Contact matrices are calculated as the preliminary features extracted from the simulations. To take advantage of the temporal information from the simulations, we enhanced contact matrices data with temporal dynamism representation, which are then modeled with unsupervised convolutional variational autoencoder (CVAE). Finally, conventional and CVAE-based clustering methods are compared with metrics to rank the submolecular structures and propose potential candidates for lead optimization. Conclusion: With no need for extensive structure-activity data, our method provides new hints for drug modification hotspots which can be used to improve drug potency and reduce the lead optimization time. It can potentially become a valuable tool for medicinal chemists.

59 BASIC BIOLOGICAL SCIENCES↗

Multisensor Agile Adaptive Sampling of Convective Storms Driven by Real-time Analytics

Convective storms vertically transport water vapor and condensate from Earth’s surface to the upper troposphere. Life on Earth is fundamentally linked to this transport which determines the hydrological cycle, and the intensity of severe weather responsible for the destruction of life and property. Despite advances in high-resolution modeling and better observational capabilities, the scientific community continues to be confronted with knowledge gaps about convective storms that limit our predictive capabilities. The ongoing developments in the high-resolution Energy Exascale Earth System Model (E3SM), large eddy simulations, and AI-based analytics to evaluate uncertainties are expected to provide a comprehensive framework for new scientific discovery. The model-experiment (MODEX) approach suggests that the aforementioned advancements in model development and AI-based inference techniques should be complemented by similar advancements in the experimental (observational) side so that the former does not outstrip the ability of the latter to provide meaningful constraints. What are the recent advancements in observations that will provide the necessary leap forward in improving our predictive capabilities? To address this question, we propose a new experimental paradigm called Multisensor Agile Adaptive Sampling (MAAS) that capitalizes on advancements in communications (5G), computational resources (edge/fog computing), sensor capabilities, and machine learning (ML) and AI techniques (Kollias et al., 2020). The MAAS framework allows for the collection of higher spatiotemporal resolution and quality observations of convective storms than is traditionally possible. The MAAS framework is scalable and applicable to atmospheric observatories such as those operated by the Department of Energy (DoE) Atmospheric Radiation Measurement (ARM) facility.

54 ENVIRONMENTAL SCIENCES↗

Simultaneous Development and Robust Optimization of a Microstructure Dependent Material

Recent microstructure characterization techniques combined with Symbolic Regression(SR)analysis has been proven to generate white box plasticity models well suited for incorporation into FEA software.The current work builds upon those efforts and demonstrates the applicability of Sequential Monte-Carlo (SMC) methods within SR analysis to condense model development and robust optimization into a single, co-dependent process. In this project, SMC methods provide a mechanism through which the observed microstructure features and associated variability can be incorporated into the discovery phase of model development and simultaneously recover approximate parameter distributions through SR analysis. The demonstration utilized a data set consisting of tensile test results from a limited number of sample specimens with corresponding EBSD data from which microstructure features were characterized.The maximum threshold stress model in the Visco-Plastic Self-Consistent (VPSC) code developed by Los Alamos National Laboratories was calibrated using mechanical test data.Synthetic volume elements with statistically equivalent microstructure were generated with DREAM3Dbased on the observed EBSD data. VPSC was used to simulate the corresponding tensile test response for each of the synthetic volume elements. The simulated microstructure and tensile test data was used astraining datafor SMC-SR algorithm and the resulting model was validated with data from the original empirical data set.

Karl Garbrecht↗