Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “reproducibility”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Reproducibility of Radiokrypton in Deep Desert Aquifers: Insights from a Decade of Research

Great technical advances have been achieved since the first atom-trap trace analysis (ATTA) -based radiokrypton application in Egypt, where 1 Myr old groundwater was discovered. Beyond advances in ATTA measurement capabilities, including reduction in sample size, analysis duration, and analytical uncertainty, major progress has been achieved over the past two decades in the sample collection and preparation techniques. These advances paved the expansion of ATTA-based noble gas applications to many other aquifers worldwide, illuminating the nature and flow pattern of deep groundwater systems. While the potential of this new analytical technique for old groundwater dating is well recognized, another important aspect yet to be examined is the reproducibility of radiokrypton in aquifers over time, i.e., how representative is a discrete groundwater sample, collected at a specific time and location, for the natural groundwater system? The likelihood of a negative answer is increased by flow-field disturbance in aquifers following massive groundwater abstraction. Here, in this work, we present repeated 81 Kr sampling and measurements in twenty-one sites over Israel, mostly of deep (up to 1 km) wells tapping confined aquifers in the arid to hyperarid Negev desert. The results demonstrate that radiokrypton measurements are indeed reproducible, even in cases where samples were collected as long as nine years apart and from highly productive (∼1 Mm 3 /yr order) pumping wells. Furthermore, many of the repeated measurements in this study (17 out of the 21 sites) were conducted with different ATTA Instruments in two different laboratories using slightly different sampling, preparation, and analysis techniques, yet with an overall good agreement. The consistency in the ATTA-based 81 Kr-dating results over time highlights the robustness of this state-of-the-art technique as a tool to unravel groundwater flow patterns and encourages further applications to many other yet-to-be-explored deep aquifers.

atom-trap trace analysis↗

Increasing the Reproducibility and Replicability of Supervised AI/ML in the Earth Systems Science by Leveraging Social Science Methods

Artificial intelligence (AI) and machine learning (ML) pose a challenge for achieving science that is both reproducible and replicable. The challenge is compounded in supervised models that depend on manually labeled training data, as they introduce additional decision-making and processes that require thorough documentation and reporting. We address these limitations by providing an approach to hand labeling training data for supervised ML that integrates quantitative content analysis (QCA)—a method from social science research. The QCA approach provides a rigorous and well-documented hand labeling procedure to improve the replicability and reproducibility of supervised ML applications in Earth systems science (ESS), as well as the ability to evaluate them. Specifically, the approach requires (a) the articulation and documentation of the exact decision-making process used for assigning hand labels in a “codebook” and (b) an empirical evaluation of the reliability” of the hand labelers. In this paper, we outline the contributions of QCA to the field, along with an overview of the general approach. We then provide a case study to further demonstrate how this framework has and can be applied when developing supervised ML models for applications in ESS. With this approach, we provide an actionable path forward for addressing ethical considerations and goals outlined by recent AGU work on ML ethics in ESS.

58 GEOSCIENCES↗

fluxfinder: An R Package for Reproducible Calculation and Initial Processing of Greenhouse Gas Fluxes From Static Chamber Measurements

Fluxes of greenhouse gases are a critical component of the earth's natural climate, but anthropogenic emissions have created an imbalance and resulted in global climate change. Quantifying the emission of these gases is vital to our understanding of their sources and sinks, both natural and anthropogenic. The static chamber method, in which a system of interest is enclosed, and gas concentrations are measured over time, is widely used to estimate fluxes of greenhouse gases. With the development of instruments such as infrared gas analyzers (IRGAs) supporting high-frequency concentration data, there is a growing need for open-source workflows to calculate fluxes. Here we present fluxfinder, an R package designed to support reproducible calculations and processing of greenhouse gas fluxes measured with the static chamber method. The package includes raw data file parsing from widely used IRGAs, metadata matching, unit conversion, flux estimations, and initial quality assurance/quality control (QA/QC). Diagnostic graphical plots provide a transparent way to differentiate between measurement issues and nonlinear behavior. The package is also designed to be easily integrated with the gasfluxes package for further fitting of nonlinear concentration-time models, allowing alternative or additional flux QA/QC. The fluxfinder package offers a flexible workflow that is easily adaptable to promote open and reproducible greenhouse gas flux estimations.

Wilson, Stephanie J.↗

Frequency reproducibility of solid-state thorium-229 nuclear clocks

Solid-state thorium-229 ( 229 Th) nuclear clocks are set to provide new opportunities for precision metrology and fundamental physics. Taking advantage of inherent low sensitivity of a nuclear transition to its environment, orders of magnitude more emitters can be hosted in a solid-state crystal compared with current optical lattice atomic clocks. Furthermore, solid-state systems needing only simple thermal control are key to the development of field-deployable compact clocks. Here we explore and characterize the frequency reproducibility of the 229 Th:CaF 2 nuclear clock transition, a key performance metric for all clocks. We measure the transition linewidth and centre frequency as a function of the doping concentration, temperature and time. We report the concentration-dependent inhomogeneous linewidth of the nuclear transition, limited by the intrinsic host crystal properties. We determine an optimal working temperature for the 229 Th:CaF 2 nuclear clock at 196(5) K, at which the first-order thermal sensitivity vanishes. This would enable in situ temperature co-sensing using different quadrupole-split lines, reducing the temperature-induced systematic shift below the 10 −18 fractional frequency uncertainty level. At 195 K, the reproducibility of the nuclear transition frequency is 220 Hz (fractionally 1.1 × 10 −13 ) for two differently doped 229 Th:CaF 2 crystals over 7 months. Furthermore, these results form the foundation for understanding, controlling and harnessing the coherent nuclear excitation of 229 Th in solid-state hosts and for their applications in constraining temporal variations of fundamental constants.

Atomic and molecular physics↗

Evaluating the factors influencing accuracy, interpretability, and reproducibility in the use of machine learning classifiers in biology to enable standardization

The complexity and variability of biological data has promoted the increased use of machine learning methods to understand processes and predict outcomes. These same features complicate reliable, reproducible, interpretable, and responsible use of such methods, resulting in questionable relevance of the derived. outcomes. Here we systematically explore challenges associated with applying machine learning to predict and understand biological processes using a well- characterized in vitro experimental system. We evaluated factors that vary while applying machine learning classifers: (1) type of biochemical signature (transcripts vs. proteins), (2) data curation methods (pre- and post-processing), and (3) choice of machine learning classifier. Using accuracy, generalizability, interpretability, and reproducibility as metrics, we found that the above factors significantly mod- ulate outcomes even within a simple model system. Our results caution against the unregulated use of machine learning methods in the biological sciences, and strongly advocate the need for data standards and validation tool-kits for such studies.

59 BASIC BIOLOGICAL SCIENCES↗

Deterministic fabrication of highly reproducible monochromatic quantum emitters in hexagonal boron nitride

Quantum emitters in hexagonal boron nitride are important room temperature single-photon sources. However, conventional fabrication methods yield quantum emitters with dispersed and inconsistent spectral profiles, limiting their potential for practical quantum applications, which demand reproducible high quality single-photon sources. Here, we report the deterministic creation of highly reproducible monochromatic quantum emitters by applying carbon-ion implantation on freestanding hexagonal boron nitride flakes, while a carbon mask with suitable thickness was adapted to optimize the implantation results. Quantum emitters fabricated using this approach exhibited thermally limited monochromaticity, with an emission center wavelength of 590.7 ± 2.7 nm, a narrow full width at half maximum of 7.1 ± 1.7 nm, an emission rate of 1 MHz without optical engineering, and exceptional stability under ambient conditions. Density functional theory calculations and scanning transmission electron microscopy suggest that these emitters are comprised of boron centered carbon tetramers. This method provides a reliable single-photon source for optical quantum computing and potential future industry-scale applications.

Hua, Muchuan [Argonne National Laboratory (ANL), A↗

Towards modelling AR Sco: calibration – reproducing high-energy pulsar emission and testing convergence to Aristotelian electrodynamics

In recent years, kinetic simulations have been crucial to further our understanding of pulsar electrodynamics. Yet, due to the large-scale separation between the gyro-period and the stellar rotation period, resolving the particle gyration has been computationally unfeasible for realistic pulsar parameters. The main aim of this work is comparing our gyro-phase-resolved model with a gyro-centric pulsar model, where our model solves the general equations of motion with included radiation reaction using a higher order numerical solver with adaptive time-steps. Specifically, we aim to (i) reproduce a pulsar’s high-energy emission maps, namely one with 10 per cent of the surface B-field strength of Vela, and the spectra produced by an independent gyro-centric pulsar emission model; and (ii) test convergence of these results to the radiation-reaction limit of Aristotelian electrodynamics. (iii) Additionally, we identify the effect that a large $E_{\parallel }$-field has on the trajectories and radiation calculations. We find that we can reproduce the curvature radiation emission maps and spectra well, using 10 per cent field strengths of the Vela pulsar and injecting our particles at a higher altitude in the magnetosphere. Using sufficiently large $E_{\parallel }$-fields, our numeric results converge to the analytic radiation-reaction limit trajectories. Additionally, we illustrate the importance of accounting for the $\mathbf {E}\times \mathbf {B}$-drift in the particle trajectories and radiation calculations, validating the Harding and collaborators’ model approach. Lastly, we found that our model deals very well with the high-radiation-reaction and high-field regimes present in pulsars.

79 ASTRONOMY AND ASTROPHYSICS↗

Experimental Benchmarking for High-Reproducibility, Cross-Institutional Evaluation of Iron Redox Electrochemistry

We present a practical case study standardizing experimental protocols between collaborators with the goal of understanding ferrous iron (Fe 2+ ) chemistry and improving the iron deposition reaction for energy-efficient, electrochemical iron production. The study of iron reactions can be difficult, as aqueous iron electrolytes exhibit complex behaviors that can lead to differing interpretation of ostensibly similar experiments. The question we want to answer: are we studying the same chemistry? Our protocols address inherent challenges such as the tendency for Fe 2+ to spontaneously oxidize to ferric iron (Fe 3+ ) and the production of hydrogen at the potentials of interest. Our standardized protocol, executed by four collaborators in different labs and institutions, yields high-reproducibility results, and identified glassy carbon electrode surface quality and Fe 3+ impurities in the salt as key factors with outsized effects on cyclic voltammetry measurements. The process of developing the protocols helped to troubleshoot underlying issues that created poorer reversibility and reproducibility. This study highlights the fact that even nominally straightforward electrochemical systems can yield vastly different outcomes due to small differences in experimental preparation and serves as a useful example for creating transparent and achievable standards for the generation of reliable datasets that can be widely used and shared.

Ketter, Benjamin [Argonne National Laboratory (ANL↗

Videos, photos, and AI-derived grain size data associated with “High-throughput AI Video Surveys Enable Reproducible Multiscale Sediment Size Mapping, with Implications for Hydrobiogeochemical Parameterization”

NOTE: The manuscript associated with this data package is currently in review. The data may be revised based on reviewer feedback. Upon manuscript acceptance, this data package will be updated with the final dataset and additional metadata. This data package is associated with the manuscript “High-throughput AI Video Surveys Enable Reproducible Multiscale Sediment Size Mapping, with Implications for Hydrobiogeochemical Parameterization” under review. This data package includes five data types: 1) raw photos and videos from drone survey and walking smartphone surveys; 2) images derived from raw videos; 3) manual labeling of reference scales; 4) metadata for all images and photo resolution derived from artificial intelligence (AI) models or manual labels, 5) grain size data obtained from AI models for all photos, 6) metadata and grain size data after quality control, 7) summaries of sample efficiency for all data, and 8) computational fluid dynamics (CFD) data used to support hydro-biogeochemical (HBGC) parameter estimation. Such data is used to 1) demonstrate significant improvements in accuracy, efficiency, and quality control for grain size data collection with the help of AI models, 2) study the spatial heterogeneity of grain size and observation reproducibility based on tens of thousands of data points generated by the AI models, and 3) evaluate the impacts of grain size heterogeneity on key HBGC parameters across sediment-to-reach and hourly-to-yearly scales. In particular, the data package contains 116 folders and 179696 files. The files include 41 videos in .mov format, 64047 photos in .jpg format, 13541 video-derived photos in .png format, 12747 segmentation mask data in .tif format, 12747 segmentation data in .json format, 24771 .csv files that with metadata and grain size for each individual photo as well as water depth and velocity data from CFD and observation, 51791 .txt files of raw AI predicted labels, and 11 flight record data in .srt format. The summary for all metadata and grain size statistics information is included in “Scales_V3_NG.csv” and “Statistics_V3_NG.csv”. The summary for data that pass data quality control (QC) level 0-2 is included in “QCStatistics_V3_NG.csv”. The QC level 0 represents photos whose photo resolution is positive, excluding photos that miss reference scale. The QC level 1 means reference scale circularity uncertainty is less than 5% for smartphone images while representing photo resolution is larger than 0.44 mm/pixel for drone images. The QC level 2 means excluding photos whose grain number is less than 100, a minimum number of grains recommended by classic literature. The summary for each video’s name, length, frame rates, survey area, grain number, survey efficiency, etc. can be found in “QCSummary_V3_NG.csv”. The summary for site name, GPS coordinates, and number of images at each site can be found in “SitesSummary_V3_*.csv” files. Overall computational efficiency summary is reported in Table 4 of accompanying manuscript. Additionally, the nitrate concentration data used in this work was downloaded from an existing dataset published on ESS-DIVE (Boat-Dragged Sensor Hanford Reach.csv; Conner A. et al., 2020). We thank the United States Forest Service, Washington Department of Fish and Wildlife, Washington Department of Natural Resources, Cowiche Canyon Conservatory, Port of Benton, and the Confederated Tribes and Bands of the Yakama Nation for access to field locations where the data were collected. We also thank the Yakama Nation Tribal Council and Yakama Nation Fisheries for working with us to facilitate data collection and optimization of data usage according to their values and worldview.

54 ENVIRONMENTAL SCIENCES↗

A reproducible study design for the MIMIC-IV in-hospital mortality task

Open, tabular electronic health record (EHR) datasets such as MIMIC-III and MIMIC-IV have become critical resources for developing machine learning (ML) models addressing clinical prediction tasks, including hospital readmission, length of stay, and in-hospital mortality (IHM). While MIMIC-III has benefited from well-established preprocessing pipelines and standardized feature sets, MIMIC-IV remains comparatively challenging to work with because there are no standardized benchmarks to support reproducibility and comparability across studies. To address this limitation, we present a rigorously curated MIMIC-IV custom feature set optimized for IHM prediction, constructed through a reproducible preprocessing pipeline and feature selection strategy.

97 MATHEMATICS AND COMPUTING↗

Enhanced climate reproducibility testing with false discovery rate correction

Simulating the Earth's climate is an important and complex problem, thus climate models are similarly complex, comprised of millions of lines of code. In order to appropriately utilize the latest computational and software infrastructure advancements in Earth system models running on modern hybrid computing architectures to improve their performance, precision, accuracy, or all three; it is important to ensure that model simulations are repeatable and robust. This introduces the need for establishing statistical or non-bit-for-bit reproducibility, since bit-for-bit reproducibility may not always be achievable. Here, we propose a short-simulation ensemble-based test for an atmosphere model to evaluate the null hypothesis that modified model results are statistically equivalent to that of the original model. We implement this test in version 2 of the US Department of Energy's Energy Exascale Earth System Model (E3SM). The test evaluates a standard set of output variables across the two simulation ensembles and uses a false discovery rate correction to account for multiple testing. The false positive rates of the test are examined using re-sampling techniques on large simulation ensembles and are found to be lower than the currently implemented bootstrapping-based testing approach in E3SM. We also evaluate the statistical power of the test using perturbed simulation ensemble suites, each with a progressively larger magnitude of change to a tuning parameter. The new test is generally found to exhibit more statistical power than the current approach, being able to detect smaller changes in parameter values with higher confidence.

Kelleher, Michael E. [Oak Ridge National Laborator↗

Reproducible emission from nonlinear random lasers

Multiple scattering of light serves as a mechanism for feedback in random lasers. Consequently, internal spatial mode patterns, lasing wavelengths, and output directionality can all be random. Strong mode interaction can occur in such devices due to spatially overlapping modes resulting in nonlinearity with respect to the pump input power. Nevertheless, temporal coherence and lasing mode amplitude can be fixed at a constant pumping rate. This is a property desirable for applications where unique randomness is exploited but expected to be reliable over time, such as physical unclonable functions. Random lasers can also be cheaply and easily fabricated, exhibit relatively low lasing thresholds and high emission intensity. However, the precise scattering properties of such structures and fluctuations in the pump field can make device emission irreproducible, thereby limiting random laser applications. Here, in this work, we directly compare the random lasing spectra from zinc oxide samples fabricated in four distinct ways: spin-coating, sputtering, solgel deposition, and atomic layer deposition. The particular method of fabrication has a strong impact. Samples made through atomic layer deposition here exhibit both reproducibility and strong nonlinearity desirable for applications. Randomness in emission spectra persists across hundreds of repeated and averaged measurements irrespective of spatial location and is demonstrably nonlinear with respect to input signal intensity.

47 OTHER INSTRUMENTATION↗

Metal Foam Morphology Affects the Run to Run Reproducibility of OER Using Nickel Catalysts

Nickel (Ni) foam-based electrodes are excellent catalysts for the oxygen evolution reaction; however, we found that the random pore-size distribution of Ni foams contributes to a significant variability in electrochemically active surface area, compromising experimental reproducibility. We provide insights into quantifying this critical material property, verified by four electrochemical laboratories.

25 ENERGY STORAGE↗

Kernel Manifolds: Nonlinear‐Augmentation Dimensionality Reduction Using Reproducing Kernel Hilbert Spaces

This paper generalizes recent advances on quadratic manifold (QM) dimensionality reduction by developing kernel methods-based nonlinear-augmentation dimensionality reduction. QMs, and more generally feature map-based nonlinear corrections, augment linear dimensionality reduction with a nonlinear correction term in the reconstruction map to overcome approximation accuracy limitations of purely linear approaches. While feature map-based approaches typically learn a least squares optimal polynomial correction term, we generalize this approach by learning an optimal nonlinear correction from a user-defined reproducing kernel Hilbert space. Our approach allows one to impose arbitrary nonlinear structure on the correction term, including polynomial structure, and includes feature map and radial basis function-based corrections as special cases. Furthermore, our method has relatively low training cost and has monotonically decreasing error as the latent space dimension increases. In conclusion, we compare our approach to proper orthogonal decomposition and several recent QM approaches on data from several example problems.

kernel methods↗

Interpretable and flexible non-intrusive reduced-order models using reproducing kernel Hilbert spaces

This paper develops an interpretable, non-intrusive reduced-order modeling technique using regularized kernel interpolation. Existing non-intrusive approaches approximate the dynamics of a reduced-order model (ROM) by solving a data-driven least-squares regression problem for low-dimensional matrix operators. Our approach instead leverages regularized kernel interpolation, which yields an optimal approximation of the ROM dynamics from a user-defined reproducing kernel Hilbert space. We show that our kernel-based approach can produce interpretable ROMs whose structure mirrors full-order model structure by embedding judiciously chosen feature maps into the kernel. The approach is flexible and allows a combination of informed structure through feature maps and closure terms via more general nonlinear terms in the kernel. We also derive a computable a posteriori error bound that combines standard error estimates for intrusive projection-based ROMs and kernel interpolants. In conclusion, the approach is demonstrated in several numerical experiments that include comparisons to operator inference using both proper orthogonal decomposition and quadratic manifold dimension reduction.

Data-driven model reduction↗

Overcoming Variability: A Reproducible Approach to SERS Detection of Nanodiamonds

Detonation nanodiamonds (DNDs) are formed at specific pressures and temperatures during explosions. Different explosives produce varied yields of DNDs within their detonation soot, with Composition B producing the highest yield. Raman spectroscopy (RS) is often used for the characterization of sp 2 - and sp 3 -hybridized carbon allotropes in carbonaceous materials because of distinct disorder and graphitic bands. Bulk diamond also gives a distinct Raman peak at 1332 cm –1 . Furthermore, as bulk diamond decreases in size to nanometer-sized species, the peak red-shifts and broadens, becoming increasingly difficult to detect with RS using visible excitations. Therefore, surface-enhanced Raman spectroscopy (SERS) was used to enhance the diamond peak of DNDs, enabling better detection and faster examination of DNDs within detonation soot. Previous literature of the SERS of DNDs delivered inconsistent results in spectral signatures and SERS substrates. Herein, refining of the methodology for the acquisition of SERS spectra of DNDs was achieved. Before any SERS experiments, the DNDs were first characterized with normal Raman (NR) and scanning electron microscopy. Two routes for SERS enhancement were evaluated: colloidal noble metal nanoparticles and evaporated silver films. Silver films produced the best signal enhancement of DNDs with the best signal-to-noise and peak enhancements observed at 20–30 nm thick silver films at 5% (∼300 μW) laser power. Consistent, reproducible SERS spectra were acquired of small aggregates of DNDs down to ∼500 nm. NR and SERS mapping analysis of DNDs before and after evaporation of silver films revealed the improvements in the detection capabilities of SERS compared with NR.

Carbon↗

Comparative Evaluation of the Ability of the MYNN‐EDMF PBL Scheme in WRF Model to Reproduce Near Surface Wind Speed Over Different Topographical Types

Abstract This study systematically evaluates the performance of the Mellor‐Yamada‐Nakanishi‐Niino‐Eddy‐Diffusion‐Mass‐Flux planetary boundary layer (PBL) scheme within the Weather Research and Forecasting (WRF) model in simulating near‐surface wind speeds across various topographies in New York State (NYS). Simulated wind speeds are compared with in‐situ measurements from 22 surface sites, grouped into six topographic categories: continental plain (CT), lakeside (LS), river valley (RV), Long Island (LI), Block Island (BI), and offshore ocean (OO). A quantitative evaluation based on Relative Euclidean Distance shows that wind speeds at the OO site are the most accurately reproduced, followed by those at LI sites, while the model performs less accurately for the remaining topographic groups. Wind speeds over CT sites tend to be overestimated by approximately 1 m/s, although their diurnal variability (DV) is well captured. In contrast, the model underestimates wind DV at LS, RV, LI, and BI sites, with the largest biases occurring at LI and BI, resulting in underestimated daytime wind speed and/or overestimated nighttime wind speed. The OO winds exhibit minimal diurnal variation, accurately captured by our WRF model. The surface wind diurnal variation is closely linked to PBL development. Among the indicators of PBL development, surface potential temperature biases most strongly correlate with wind speed biases. Our WRF model faces challenges in capturing the distinctions between winds influenced by local circulations and those over continental plains, and the significantly stronger winds at OO compared to BI. Potential causes for these biases are discussed, offering pathways for improving surface wind simulations in future.

54 ENVIRONMENTAL SCIENCES↗

Design-driven optimization of low-cost reagent formulations for reproducible and high-yielding cell-free gene expression

Access to recombinant proteins is vital in basic science and biotechnology research. Cell-free gene expression systems provide one approach to address this need, but widespread utilization remains limited by the cost, complexity, and inconsistency of current platforms. To address these limitations, we carry out a multi-dimensional definitive screening design to reduce the number of reagent components and remove costly secondary energy substrates. From 1,231 different reagent formulations, we discover a simple and reproducible system based on 12 components. The optimized reagent formulation can produce 2.4 ± 0.3 g/L of protein product at the 15-µL scale (~$\$60$/gprotein) and 3.7 ± 0.2 g/L (~$\$39$/gprotein) at the 4-mL scale with oxygen supplementation. This provides an average 95% reduction in cost over previous cell-free reagent formulations. We further show that the optimized reagent formulation can produce nucleoside triphosphates from nitrogenous bases and ribose and that it is robust to failure across batches of cell lysates, users/locations, and in the synthesis of more than 20 different proteins. For example, we demonstrate the production of fifteen therapeutically relevant products, including full-length aglycosylated monoclonal antibodies. We anticipate that our optimized reagent formulation will democratize the use of cell-free systems for protein manufacturing and synthetic biology applications.

Biologics↗