Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “trustworthiness”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

Beyond Fair: Engagement, Data Usability, and Open Community Productivity through the NASA Open Science Data Repository

The FAIR principle (findable, accessible, interoperable, and reusable) governs the storage and sharing of NASA space biology and health data[1]. These guiding principles maximize reuse of data and the reproducibility of scientific findings. The NASA Open Science Data Repository (OSDR; an expansion of NASA GeneLab) was built on the FAIR principles and houses over 500 studies and close to 1000 datasets from decades of space life sciences experiments. OSDR embodies the FAIR principles through data governance that includes mediated, embargoed, and fully open access data. The FAIR data governance principles were recently proposed to be expanded to encompass a FAIREST framework for assessing research data repositories (FAIR + Engagement, Social connections, and Trust)[2]. FAIREST emphasizes the importance of data repositories engaging with the scientific community and gaining the trust of researchers regarding data quality. Trust also refers to the TRUST principles developed for assessment of digital repositories: Transparency, Responsibility, User Focus, Sustainability, Technology[3]. We present the “Open Science for Life in Space” Analysis Working Groups (AWGs) as evidence regarding the power of engagement, social connections, and trust which has enhanced OSDR’s capabilities and productivity. AWG members engage in two main activities. One, members provide feedback on OSDR scientific standards for data ingestion, curation, and reuse (study, subject and assay metadata; processing pipelines; dataset formats and uniformed structures for machine-readability). Two, AWG members collaborate to mine-reuse OSDR data to conduct scientific analysis. With nearly 800 active members, the AWGs have resulted in 32 publications re-using OSDR data and contributed many papers in two major special issues in Cell (2020) and Nature (2024). AWGs also serve as networking groups, facilitate social connections between researchers at all levels of experience, and also have a social online ‘Forum’ used to keep members informed on projects and opportunities. This community-centric, productive, and trustworthy data culture has resulted in a broader effect with international space agencies, academics, and the commercial space sector wanting to submit their data to OSDR. Ten studies of Inspiration 4 data were recently publicly released by OSDR, as were some JAXA human data. Coming up soon in OSDR are data submissions from the European Space Agency, Virgin Galactic PIs, and SpaceX Polaris Dawn. A major benefit of OSDR is the array of standardized and uniformly formatted data (which was developed through AWG member consensus), from which visualization tools, analysis tools, and machine learning models can be built or trained. This talk will cover the Multi-Study Visualization Tool, the Environmental Data Application, RadLab, and a UCSF-NSF funded knowledge graph biomedical health discovery tool ‘SPOKE’ currently being integrated with OSDR. OSDR also provides training programs in bioinformatics and machine learning to improve the scientific community’s awareness of data availability and to boost their ability to perform data analysis. The increasing engagement of the scientific community and the public with technologies powered by artificial intelligence (AI) heightens the need for data analysis to be transparent. The AI for Life in Space initiative leverages the data products provided in OSDR to train AI models, with an emphasis on explainable and trustworthy AI, which would not be possible without FAIR data and metadata. Overall, here we will demonstrate the importance for NASA life sciences data repositories to adhere to the FAIREST framework, by providing examples and success stories from different aspects of OSDR.

data↗

Enabling Scale-Up Through Multi-Fidelity Adaptive Computing

We present ideas from our ongoing work in adaptive computing - an optimization framework that allows us to strategically deploy various fidelity level experiments and simulations to guide decision making. The framework aims to enable uncertainty quantified scale-up of simulations and experiments, which causes increased complexity. A key feature of the framework is the integration of user-specified local model trustworthiness estimates. Adaptive sampling strategies allow us to optimally exploit the multiple fidelity level information and trustworthiness measures to arrive at the best decisions within a highly limited budget of objective function evaluations.

adaptive sampling↗

Sustainable Aviation Operations and the Role of Information Technology and Data Science: Background, Current Status and Future Directions

This paper reviews the achievements of the international community towards environmentally friendly aviation operations, also referred to as Sustainable Aviation Operations in the last 25 years and the aspirations and goals to limit the impact of aviation and climate in the future. The framework for achieving global progress is provided by the International Civil Aviation Organization. NASA and FAA supported research and development to advance ATM concepts, and implemented the technology, concepts, and procedures that were responsible for creating fuel efficient flights. Historically aviation operations have been analyzed using physics-based models and provide information for making operational decisions. Future developments in aviation operations require new concepts, procedure, modeling, and analysis techniques. There is an increasing interest in applying methods based on Machine Learning Techniques to problems in Air Traffic Management. Aviation operations involving many decision makers, multiple objectives, poor or unavailable physics-based models and the availability of a rich historical database provide opportunities to exploit the richness of data-driven methods. The promises and challenges in applying Machine Learning Techniques to Air Traffic Management are discussed in the paper along with the testing and trustworthiness required for adoption of the techniques in operations.

Sustainable Aviation, Data Science, Machine Learni↗

Sustainable Aviation Operations and the Role of Information Technology and Data Science: Background, Current Status and Future Directions

This paper reviews the achievements of the international community towards environmentally friendly aviation operations, also referred to as Sustainable Aviation Operations in the last 25 years and the aspirations and goals to limit the impact of aviation and climate in the future. The framework for achieving global progress is provided by the International Civil Aviation Organization. NASA and FAA supported research and development to advance ATM concepts, and implemented the technology, concepts, and procedures that were responsible for creating fuel efficient flights. Historically aviation operations have been analyzed using physics-based models and provide information for making operational decisions. Future developments in aviation operations require new concepts, procedure, modeling, and analysis techniques. There is an increasing interest in applying methods based on Machine Learning Techniques to problems in Air Traffic Management. Aviation operations involving many decision makers, multiple objectives, poor or unavailable physics-based models and the availability of a rich historical database provide opportunities to exploit the richness of data-driven methods. The promises and challenges in applying Machine Learning Techniques to Air Traffic Management are discussed in the paper along with the testing and trustworthiness required for adoption of the techniques in operations.

Sustainable Aviation, Data Science, Machine Learni↗

Trusted Simulation: Considering Model Quality in the Context of User Trust

A high‐quality simulation model should help its users to easily and appropriately calibrate their trust in the model. Traditional evaluation metrics such as validation and robustness are necessary but insufficient for this task. Trust calibration depends on factors like the model's transparency, applicability to intended use, usability, reputation, and consideration of potential bias. This article proposes a framework for designing and evaluating system dynamics models by considering factors that contribute to the proper calibration of user trust. This framework takes inspiration from trusted artificial intelligence, broadening our traditional concept of model quality and explicitly focusing on what users need to consider a model trustworthy and to understand the model's relevance to its intended purpose. The trusted simulation framework can improve our integration of model quality activities throughout the modeling process, leading to more impactful and better‐targeted model design, development, and evaluation.

Naugle, Asmeret Bier [Sandia National Laboratories↗

Enabling end-to-end secure federated learning in biomedical research on heterogeneous computing environments with APPFLx

Facilitating large-scale, cross-institutional collaboration in biomedical machine learning (ML) projects requires a trustworthy and resilient federated learning (FL) environment to ensure that sensitive information such as protected health information is kept confidential. Specifically designed for this purpose, this work introduces APPFLx - a low-code, easy-to-use FL framework that enables easy setup, configuration, and running of FL experiments. APPFLx removes administrative boundaries of research organizations and healthcare systems while providing secure end-to-end communication, privacy-preserving functionality, and identity management. Furthermore, it is completely agnostic to the underlying computational infrastructure of participating clients, allowing an instantaneous deployment of this framework into existing computing infrastructures. Experimentally, the utility of APPFLx is demonstrated in two case studies: (1) predicting participant age from electrocardiogram (ECG) waveforms, and (2) detecting COVID-19 disease from chest radiographs. Here, ML models were securely trained across heterogeneous computing resources, including a combination of on-premise high-performance computing and cloud computing facilities. By securely unlocking data from multiple sources for training without directly sharing it, these FL models enhance generalizability and performance compared to centralized training models while ensuring data remains protected. In conclusion, APPFLx demonstrated itself as an easy-to-use framework for accelerating biomedical studies across organizations and healthcare systems on large datasets while maintaining the protection of private medical data.

Biomedical Research↗

Resilient information and inference networks under mixed-trust sensing

With ubiquitous digitization, sensing, and computational intelligence deployed in increasingly more and broader domains, including critical infrastructure, potentially misleading and destabilizing effects of multimodal anomalies and adversarial behavior are growing in importance. Here, we develop randomized and reinforcement learning-based strategies for strategically recruiting and utilizing deployed (and, thus, vulnerable and potentially faulty and/or compromised) nodes from information and inference networks, while defending against adversaries that attempt to misguide assessments of inferred variables. Recognizing that, besides communication and other costs, sampling from any observable node can either provide true data or dangerously expose our inference to misinformation (without being easily distinguishable what actually happens), the proposed strategies proceed by progressively recruiting nodes and cautiously scaling their information contribution based on assumed, or, in our reinforcement learning approach, intelligently weighed trustworthiness, with the learning approach also considering network-wide, threat-inclusive risk/value tradeoffs. While avoiding the hardware, communication, analytical and computational burden of explicit redundancy, the proposed defensive schemes enable on-the-fly assessments of underlying processes, and system-wide situational awareness with demonstrable resilience against adversarial activities.

97 - MATHEMATICS AND COMPUTING↗

AI-assisted object condensation clustering for calorimeter shower reconstruction at CLAS12

Several nuclear physics studies using the CLAS12 detector rely on the accurate reconstruction of neutrons and photons from its forward angle calorimeter system. These studies often place restrictive cuts when measuring neutral particles due to an overabundance of false clusters created by the existing calorimeter reconstruction software. In this work, we present a new AI approach to clustering CLAS12 calorimeter hits based on the object condensation framework. The model learns a latent representation of the full detector topology using GravNet layers, serving as the positional encoding for an event’s calorimeter hits which are processed by a Transformer encoder. This unique structure allows the model to contextualize local and long range information, improving its performance. Evaluated on one million simulated $e^-$ $+$ $p$ collision events, our method significantly improves cluster trustworthiness: the fraction of reliable neutron clusters, increasing from 8.88% to 30.73%, and photon clusters, increasing from 51.07% to 64.73%. In conclusion, our study also marks the first application of AI clustering techniques for hodoscopic detectors, showing potential for usage in many other experiments.

Calorimeters↗

Debunking common myths in coastal circulation modeling

Despite tremendous progress in algorithm development, computational efficiency and transition into operations over the past two decades, coastal modeling still lacks scientific rigor due to proliferation of many ‘gray’ areas related to various modeling choices made by modelers. Here, in this paper, we propose some guiding principles for the modeling community to improve performance, and we also debunk commonly held myths that make the coastal modeling lack rigor. Using our own experience in developing seamless cross-scale unstructured-grid based models for the past two decades, we describe in unprecedented detail the end-to-end modeling process (i.e., from digital elevation models (DEMs) to mesh generation to post analysis), and demonstrate that defensible modeling is within reach for any end user by following three guiding principles: (1) Bathymetry is a first order forcing in coastal domains and thus should be respected in all aspects of modeling; (2) Oceanographic processes are driven across multiple spatial scales and so models should enable appropriate resolution as needed; and (3) Model assessment should focus on physical processes. Through qualitative and quantitative model assessments, we demonstrate the fundamental role played by bathymetry/topography as embedded in DEMs in making the results defensible, which is unfortunately glossed over in many modeling studies. Focusing on process-based assessment simplifies the calibration process. A major conclusion of this work is that model developers and operators should maximize the scientific rigor for in silico oceanography by avoiding some common pitfalls that rely on error compensation at the expense of representation of physical system processes. We present some best practice procedures for defensive and trustworthy numerical modeling.

54 ENVIRONMENTAL SCIENCES↗

Advanced surrogate model for electron-scale turbulence in tokamak pedestals

We derive an advanced surrogate model for predicting turbulent transport at the edge of tokamaks driven by electron temperature gradient (ETG) modes. Our derivation is based on a recently developed sensitivity-driven sparse grid interpolation approach for uncertainty quantification and sensitivity analysis at scale, which informs the set of parameters that define the surrogate model as a scaling law. Our model reveals that ETG-driven electron heat flux is influenced by the safety factor q, electron beta β e and normalized electron Debye length λ D , in addition to well-established parameters such as the electron temperature and density gradients. To assess the trustworthiness of our model's predictions beyond training, we compute prediction intervals using bootstrapping. The surrogate model's predictive power is tested across a wide range of parameter values, including within-distribution testing parameters (to verify our model) as well as out-of-bounds and out-of-distribution testing (to validate the proposed model). Overall, validation efforts show that our model competes well with, or can even outperform, existing scaling laws in predicting ETG-driven transport.

fusion plasma↗

Machine learning-accelerated discovery of heat-resistant polysulfates for electrostatic energy storage

The development of heat-resistant dielectric polymers that withstand intense electric fields at high temperatures is critical for electrification. Balancing thermal stability and electrical insulation, however, is exceptionally challenging as these properties are often inversely correlated. A traditional intuition-driven polymer design approach results in a slow discovery loop that limits breakthroughs. Here we present a machine learning-driven strategy to rapidly identify high-performance, heat-resistant polymers. A trustworthy feed-forward neural network is trained to predict key proxy parameters and down select polymer candidates from a library of nearly 50,000 polysulfates. The highly efficient and modular sulfur fluoride exchange click chemistry enables successful synthesis and validation of selected candidates. A polysulfate featuring a 9,9-di(naphthalene)-fluorene repeat unit exhibits excellent thermal resilience and achieves ultrahigh discharged energy density with over 90% efficiency at 200 °C. Its exceptional cycling stability underscores its promise for applications in demanding electrified environments.

Li, He↗

AutoLabs: cognitive multi-agent systems with self-correction for autonomous chemical experimentation

The automation of chemical research through self-driving laboratories (SDLs) promises to accelerate scientific discovery, yet the reliability and granular performance of the underlying AI agents remain critical, under-examined challenges. In this work, we introduce AutoLabs, a self-correcting, multi-agent architecture designed to autonomously translate natural-language instructions into executable protocols for a high-throughput liquid handler. The system engages users in dialogue, decomposes experimental goals into discrete tasks for specialized agents, performs tool-assisted stoichiometric calculations, and iteratively self-corrects its output before generating a hardware-ready file. We present a comprehensive evaluation framework featuring five benchmark experiments of increasing complexity, from simple sample preparation to multi-plate timed syntheses. Through a systematic ablation study of 20 agent configurations, we assess the impact of reasoning capacity, architectural design (single- vs. multi-agent), tool use, and self-correction mechanisms. Our results demonstrate that agent reasoning capacity is the most critical factor for success, reducing quantitative errors in chemical amounts (nRMSE) by over 85% in complex tasks. When combined with a multi-agent architecture and iterative self-correction, AutoLabs approaches expert-authored reference procedures on the benchmark (F1-score > 0.89) on challenging multi-plate syntheses. These findings establish a clear blueprint for developing robust and trustworthy AI partners for autonomous laboratories, highlighting the synergistic effects of modular design, advanced reasoning, and self-correction to ensure both performance and reliability in high-stakes scientific applications. Code: https://github.com/pnnl/autolabs

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Uncertainty quantification for molecular property predictions with graph neural architecture search

Graph Neural Networks (GNNs) have emerged as a prominent class of data-driven methods for molecular property prediction. However, a key limitation of typical GNN models is their inability to quantify uncertainties in the predictions. This capability is crucial for ensuring the trustworthy use and deployment of models in downstream tasks. To that end, we introduce AutoGNNUQ, an automated uncertainty quantification (UQ) approach for molecular property prediction. AutoGNNUQ leverages architecture search to generate an ensemble of high-performing GNNs, enabling the estimation of predictive uncertainties. Our approach employs variance decomposition to separate data (aleatoric) and model (epistemic) uncertainties, providing valuable insights for reducing them. In our computational experiments, we demonstrate that AutoGNNUQ outperforms existing UQ methods in terms of both prediction accuracy and UQ performance on multiple benchmark datasets, and generalizes well to out-of-distribution datasets. Additionally, we utilize t-SNE visualization to explore correlations between molecular features and uncertainty, offering insight for dataset improvement. AutoGNNUQ has broad applicability in domains such as drug discovery and materials science, where accurate uncertainty quantification is crucial for decision-making.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Experimental uncertainty quantification using templates of expected measurement uncertainties for fast neutron-induced total, capture, and scattering cross sections

Careful experimental uncertainty quantification (UQ) is key for developing trustworthy evaluated nuclear data. Templates to account for missing or under-reported experimental uncertainties were recently developed by the covariance committee of Cross Section Evaluation Working Group (CSEWG). In this work, we illustrate the practical application and limitations of these templates for selected neutron-induced reactions, including (n, tot), (n, γ), and (n, xn) in the fast energy range, to illustrate their use in data analyses for nuclear data evaluations. We show that while the templates provide consistent framework, proper implementation still requires detailed knowledge of experimental conditions and careful treatment of nonlinear effects in cross section derivation. Case studies highlight how template-assisted UQ improves consistency with previous evaluations such as ENDF/B and reveals open challenges in propagating uncertainties across different energy regimes. The main contribution of this paper is to connect formal template recommendations with their use in practical evaluation workflows, clarifying both their benefits and current limitations.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Dynamic in-context learning with conversational models for data extraction and materials property prediction

The advent of natural language processing and large language models (LLMs) has revolutionized the extraction of data from unstructured scholarly papers. However, ensuring data trustworthiness remains a significant challenge. In this paper, we introduce PropertyExtractor, an open-source tool that leverages advanced conversational LLMs such as Google gemini-pro and OpenAI gpt-4, blends zero-shot with few-shot in-context learning, and employs engineered prompts for the dynamic refinement of structured information hierarchies—enabling autonomous, efficient, scalable, and accurate identification, extraction, and verification of material property data. Our tests on material data demonstrate precision and recall that exceed 95% with an error rate of ∼9%, highlighting the effectiveness and versatility of the toolkit. Finally, databases for 2D material thicknesses, a critical parameter for device integration, and energy bandgap values are developed using PropertyExtractor. In particular, for the thickness database, the rapid evolution of the field has outpaced both experimental measurements and computational methods, creating a significant data gap. Our work addresses this gap and showcases the potential of PropertyExtractor as a reliable and efficient tool for the autonomous generation of various material property databases, advancing the field.

Ekuma, Chinedu E. (ORCID:0000000258527556)↗

Reference Correlations for the Density and Viscosity of Molten Alkali and Alkaline Earth Fluoride Salts

While there is a significant body of literature pertaining to thermophysical property measurements of molten salts, there is often a wide degree of variability among independent measurements of the same compounds. As such, the scientific community benefits greatly from an unbiased, independent assessment of duplicate datasets, so that reference correlations which describe these thermophysical properties as functions of temperature can be determined and then commonly used by researchers, scientists, and engineers. With regard to molten fluoride compounds, a significant time has elapsed since density and viscosity reference correlations have been determined; Janz conducted the most recent effort, in 1988, to provide reference correlations for the densities and viscosities of molten fluoride compounds via the National Standard Reference Data System coordinated by the National Bureau of Standards. Since then, new data have been published for molten fluoride compounds, and a new precedent has surfaced for putting forth reference correlations that involve fitting to multiple primary datasets. In this work, reference correlations are put forth for molten alkali and alkaline earth fluoride compounds in an effort to provide updated, improved correlations for general use. For molten alkali fluoride densities, estimated uncertainties with a 95% confidence interval are summarized as follows: LiF (0.63%), NaF (0.48%), KF (0.76%), RbF (0.93%), and CsF (0.75%). For molten alkaline earth fluoride densities, an estimated uncertainty was not able to be quantified for BeF 2 because of limited data; however, estimated uncertainties with a 95% confidence interval are summarized as follows for the remaining alkaline earth fluorides: MgF 2 (1.5%), CaF 2 (0.92%), SrF 2 (1.6%), and BaF 2 (0.23%). For molten alkali fluoride viscosities, uncertainty was not able to be quantified for RbF and CsF because of limited data; however, estimated uncertainties with a 95% confidence interval are summarized as follows for the remaining alkali fluorides: LiF (4.4%), NaF (3.0%), and KF (4.0%). For molten alkaline earth fluoride viscosities, limited consistent data resulted in the recommendation of single datasets (from literature) that are deemed to be the most trustworthy based on the quality of the underlying experimental studies.

Birri, A. [Oak Ridge National Laboratory (ORNL), O↗

Interpreting AI for fusion: An application to plasma profile analysis for tearing mode stability

Artificial intelligence models have demonstrated strong predictive capabilities for various instabilities in fusion devices such as Tokamaks, including tearing modes (TM), edge localized modes, and disruptive events, but their opaque nature raises concerns about safety and trustworthiness when applied to fusion power plants. Here, we present a physics-based interpretation framework using a TM prediction model as a demonstration that is validated through a dedicated DIII-D TM avoidance experiment. By applying Shapley analysis, we identify how profiles such as rotation, temperature, and density contribute to the model's prediction of TM stability. Our analysis shows that in our experimental scenario, core electron temperature and rotation peaking play the primary role in TM stability, while density changes have smaller effects on stability. We show that off-axis ion temperature stabilizes TMs, suggesting that off-axis neutral beam heating can further stabilize this scenario. This work presents a generalizable ML-based event prediction methodology, from training to physics-driven interpretation, bridging the gap between physics understanding and opaque ML models.

Farre-Kaga, Hiro J. [Princeton Univ., NJ (United S↗