Engineering PapersSearch

SEARCH · Engineering Papers

Results for “data normalization”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Measurements of Normalized Differential Cross Sections of Inclusive η Production in e + e − Annihilation at Energy from 2.0000 to 3.6710 GeV

Using data samples collected with the BESIII detector operating at the BEPCII storage ring, the cross section of the inclusive process e + e − → η + X , normalized by the total cross section of e + e − → hadrons , is measured at eight center-of-mass energy points from 2.0000 to 3.6710 GeV. These are the first measurements with momentum dependence in this energy region. Our measurement shows a significant discrepancy compared to the existing fragmentation functions. To address this discrepancy, a new QCD analysis is performed at the next-to-next-to-leading order with hadron mass corrections and higher twist effects, which can explain both the established high-energy data and our measurements reasonably well. Published by the American Physical Society 2024

Physics

Alternating conduction and convection drying of paper – an experimental analysis with a continuous data acquisition approach

In conventional multi-cylinder drying of paper and board, both conductive drying from steam-heated dryer cylinders and convective drying by flowing air over the paper surface in the pockets are used. Conductive drying from steam-heated drying cylinders is a critical component in providing the necessary thermal energy to paper and board as they dry. Steam temperature and internal and external resistances at the contacting surface are critical process parameters influencing the conductive drying process. An experimental setup was developed to study the alternating conductive and convective drying of paper and board. Paper sheet moisture, temperature, and temperature distribution within the heated platen and the instantaneous heat flux as the sheet was being dried were measured. The instantaneous heat flux, contact heat transfer coefficient, and drying rates were determined as drying proceeds. Experimental results, as well as comparisons to literature and commercial data, are presented. The conductive heat transfer coefficients determined were compared to traditional correlations normally used in the modeling of paper drying. Similarly, the convective heat and mass transfer coefficients are also determined and compared to literature data. In addition to the evaluation of alternating conductive and convective drying characteristics of paper and board, the potential inclusion of auxiliary energy components will also be included. Experimental results from the conduction and convection drying system are presented. Furthermore, this data will be useful in process development, intensification of manufacturing processes, and modeling and simulation of paper drying processes.

42 ENGINEERING

Determining reference standard strength for neutron-irradiated reduced activation ferritic/martensitic steel F82H by Bayesian method

The deterministic approach widely adopted in the design of structural components relies on systematically defined design limits using empirically determined safety factors. However, this approach is not always appropriate because structures are subjected to a variety of loads in the practical environment, which may result in excessively conservative design limits. In recent years, a more rigorous probabilistic approach that incorporates material strength distributions has become an important solution. In the probabilistic approach, the probability density functions of material strength properties underpin the design criteria. Here, the objective of this study is to identify the density distribution functions that best describe tensile properties of irradiated F82H to define a reference strength for DEMO design. Due to the limited number of existing data, this study specifically employs a Bayesian prediction method based on Monte Carlo simulations to determine a material reference value with statistical reliability and to investigate its effectiveness. For example, the dependence of tensile properties of 300 °C irradiated materials on irradiation damage and the range predicted by 95% Bayesian estimation was evaluated. As a statistical model for the dose dependence of statistical parameters, the normal distribution exhibited a better fit for 0.2% proof strength and tensile strength, whereas the distribution of total elongation data gave comparable reference values for both the normal and Weibull distribution models. Both models gave comparable criteria for the distribution of total elongation data. The Weibull model also gave better results for uniform elongation. The function best describing the model was a logarithmic law for both 0.2% proof strength and tensile strength, while a power law for both total and uniform elongation, which allowed for more comprehensive data prediction of irradiation data with statistical accuracy for DEMO reactor design.

36 MATERIALS SCIENCE

Outdoor Deployment Data for a Four-Terminal GaAs//Si Tandem Solar Mini-Module

This dataset contains the complete outdoor measurement and analysis data for a mechanically stacked, four-terminal (4T) gallium arsenide (GaAs)//silicon (Si) tandem solar mini-module deployed from October 2019 to January 2021 at the Solar Radiation Research Laboratory (SRRL) in Golden, Colorado, USA. The data support a performance modeling and degradation analysis framework for tandem photovoltaic devices, as described in the accompanying publication. The dataset includes: (1) current–voltage (J–V) characteristics of each sub-cell measured approximately every five minutes, with extracted performance parameters; (2) spectral irradiance from an EKO MS-710 WISER spectroradiometer, along with derived spectral mismatch ratios (SMR) and average photon energy (APE); (3) one-minute resolution meteorological data from the co-located SRRL weather station and GPS-derived precipitable water vapor (PWV); (4) pre-deployment laboratory characterization (external quantum efficiency, J–V curves, standard test conditions parameters); (5) outdoor-extracted temperature and PWV correction coefficients; and (6) PVcircuit equivalent-circuit simulation outputs used for model validation. Degradation rates of −4.1 ± 0.2 %/year (GaAs) and −2.5 ± 0.9 %/year (Si) were determined using a filtering and normalization methodology adapted for fixed-tilt tandem modules. All data are provided in open, portable formats (Apache Parquet, CSV, JSON) to enable full reproducibility of the published analysis.

14 SOLAR ENERGY

Data‐Efficient Generation of Synthetic Microstructures of Polymer‐Bonded Energetic Material With Fine‐Tuned Stable Diffusion

Among current deep learning approaches for synthetic image generation, diffusion-based models stand out in terms of algorithmic stability and ability to retain high-fidelity image features with detailed resolution. Here, in this work, we employ Dreambooth, a method for fine-tuning Stable Diffusion, on X-ray CT images of microstructure of the polymer-bonded form (PBX) of a commonly used high explosive, Pentaerythritol tetranitrate (PETN), which yields generative models for creating synthetic PBX images. The models developed here represent five classes (or ‘lots’) of microstructures and demonstrate successful generation of images of each class with high fidelity, as verified by computed classification accuracy of ∼ 94% or higher. Data augmentation afforded by such image synthesis can be used to more reliably decipher underlying statistics, build processing-structure correlations, recognize off-normal structural anomalies, and identify age-related changes. Ideas related to converting image data into appropriate density mapping and performing mesoscale simulation or surrogate modeling of detonation are also discussed.

Dreambooth

Multi-Objective Boundary Analysis of Discrete and Integrated SiC FET Modular Non-inverting Buck and Boost Converters for Fuel Cell EVs

This paper presents a multi-objective analysis of discrete and integrated SiC FET-based non-inverting buck-boost converter modules for modular fuel cell electric vehicle (EV) systems. Two converter ratings, 60 kW and 90 kW, are evaluated for both implementations, scalable up to 420 kW and 450 kW, respectively. Performance is assessed across efficiency, volumetric and gravimetric power density, cost, thermal stress, and estimated lifetime, where lifetime is derived from SiC FET B10 power-cycling data and junction temperature variations at rated power. A normalized overall performance index combined with a Pareto-boundary framework is used to identify configurations that optimally balance competing objectives. Results show that most configurations lie on the Pareto front, providing balanced trade-offs, while certain high-power discrete (90 kW at 450 kW) and integrated (60 kW at 180−420 kW) configurations are dominated. In general, discrete modules are more favorable for lower-power modular systems due to higher power density and lower cost, whereas integrated modules become more advantageous at higher power levels due to improved thermal behavior and longer lifetime. These findings provide practical design guidance for scalable fuel cell converter architectures and highlight the importance of system-level trade-offs in modular power electronics design.

Asa, Erdem [ORNL] (ORCID:0000000190884812)

Unbinned extraction of $γ$ from $B\to DK$ with normalizing flows

We introduce an unbinned method for extracting the CKM angle $γ$ from the decay chain $B^\pm \to (D \to K_S π^+ π^-) K^\pm$ using normalizing flows (NFs). The NFs, trained on $D$ decay data, learn a faithful continuous representation of the amplitude and strong phase variation over the $D\to K_Sπ^+π^-$ Dalitz plot whose fidelity improves with increased data sample sizes. With this input, the $B$ decay data can be used to extract the parameters $r_B$, $δ_B$, and $γ$. We test the method on Monte Carlo generated data, where it successfully recovers the injected value of $γ$ within uncertainties. The present implementation propagates statistical uncertainties from finite training data via an ensemble of independently trained flows, and does not attempt to capture the effects of systematic experimental errors. We explore two versions of the method that differ in how the trigonometric constraint on phase variation is encoded, and comment on the possible extension to Bayesian NFs, which would provide direct uncertainty estimates on the learned densities without requiring ensemble training.

Grossman, Yuval [Cornell U., LEPP]

Improving the National Solar Radiation Database (NSRDB) Using a Physics-Based Direct Normal Irradiance (DNI) Model

The National Solar Radiation Database (NSRDB) is a widely used resource providing satellite-derived solar data across the United States and globally. While the NSRDB employs a physical model for computing global horizontal irradiance (GHI), its current method for estimating cloudy-sky direct normal irradiance (DNI) relies on surface observations and empirical models. Recently, a novel physics-based approach, the Fast All-Sky Radiation Model for Solar applications with DNI (FARMS-DNI), was developed to enhance the DNI forecasting. FARMS-DNI incorporates both direct and scattered solar radiation within the circumsolar region, resulting in improved day-ahead DNI predictions when integrated into the Weather Research and Forecasting model with Solar extensions (WRF-Solar). This study integrates FARMS-DNI into the NSRDB algorithm to generate high-resolution DNI data from satellite resources. Our findings reveal that FARMS-DNI effectively mitigates the substantial DNI overestimation present in the conventional NSRDB across surface sites, particularly in conditions categorized as cloudy overcast. Consequently, this innovative model substantially enhances the overall accuracy of the NSRDB.

Xie, Yu

CONTROL AND DATA ACQUISITION IN A CYBER-PHYSICAL MIDSTREAM TESTBED

This thesis presents the development of a laboratory-scale cyber–physical midstream pipeline testbed designed to address this gap and support research in industrial control systems security. The platform integrates pumps, valves, sensors, programmable logic controllers (PLCs), and a human–machine interface (HMI) to emulate the monitoring and control architecture of real pipeline operations. The physical process is implemented as a closed-loop liquid circulation system designed to replicate flow behavior characteristic of midstream pipeline infrastructure. The testbed enables real-time data acquisition of key process variables, including flow rate and pressure facilitating the generation of datasets representative of normal pipeline operation. A threat model encompassing common ICS attack vectors was developed, including sensor spoofing, command injection, false data injection, denial-of-service attacks, and relay manipulation. Multiple attack scenarios were implemented and evaluated to demonstrate how cyber intrusions targeting sensors, actuators, networks, and software propagate into measurable physical consequences in pipeline flow and pressure. The developed platform serves as a practical, cost-effective environment for experimentation, education, and future cybersecurity research in midstream pipeline systems.

42 ENGINEERING

SSTDR and FDR Detection of Un-Energized and Energized Cable Anomalies Including Thermal Degradation Using Machine Learning

Historically, cables are initially qualified for nuclear power plant use for 40 years. As plants extend their operating license to 60 and 80 years, continued use of these cables must shift to a performance-based approach since it is cost prohibitive to completely replace cables that are likely still capable of performing their design function. A variety of cable tests are available and are commonly applied during outages when the cables can be taken out of service. Frequency domain reflectometry (FDR) is one of these test methods that is being more broadly accepted and used because it not only detects anomalies along the cable with a low-voltage signal that does not stress the cable insulation, but the technique also locates the anomalies. This supports follow-up local inspection and local repair or partial replacement of a damaged cable segment. Currently, FDR testing is only applied to cables that are taken out of service since the test instrument would be damaged by operational voltages. A related technology that has found some acceptance in the aircraft and rail industry is spread spectrum time domain reflectometry (SSTDR). This technology has been implemented with a custom commercial instrument by LiveWire Innovation that is designed to operate on live cables up to 1000 volts and with a bandwidth of 48 MHz. Initial evaluation by the Pacific Northwest National Laboratory (PNNL) of the Live Wire system indicated that a broader bandwidth (BW) SSTDR may be better for many kinds of flaws. This led PNNL to develop an SSTDR laboratory instrument suitable for tests up to 500 MHz bandwidth. Testing on energized cables is also desirable for online monitoring systems so an inductive clamshell coupler was developed that allows energized cables to be tested up to at least 5 kV and likely higher voltage levels. Dielectric spectroscopy and tan delta testing plus various laboratory destructive tests were included in this data acquisition campaign directed to feed a machine learning (ML) study. With these kinds of developments, online energized cable tests may be possible with industrial adoption of such hardware advances but it will be completely impractical to have highly skilled data analysts continually examine these complex signals for indications of damage or compromised conditions. If online testing is to be implemented in new test hardware, it must be accompanied by software that can interpret the signals and alert plant operators of changing or degraded conditions. The thermally aged, shielded cable investigated here was separately treated for ML analysis. Visual analysis of electrical data showed generally increasing peaks where the cable entered and exited the oven. These peaks were not exactly aligned with expected locations, but these differences were attributed to velocity of propagation calibration errors. Only supervised ML was applied to the thermally aged data as this data was only available shortly before the committed publication date of this report. The supervised ML was structured to divide the 0 to 70-day responses as ‘normal’ from 0 to 35 days or ‘anomalous’ from 36 to 70 days, based on cable tensile elongation at break (EAB) insulation characterization. Using 80% of the data for training and 20% for testing, the supervised ML predicted normal versus anomalous was 70% accurate. Important conclusions include: • Accuracy to predict the presence of cable damage is improved from the 2023 effort by more training data. Weighted accuracies for comparisons among the instruments ranged from 67 to 89 % for unsupervised ML and 71 to 99% for supervised ML. • Based on the synthetic data tests, the unsupervised models are more generalizable to unseen anomalies. The Multi-Layer Perceptron classifier (MLP) model reported as high as 99.7% accuracy on the test data, but this dropped to 58.3% when tested on the synthetic data. In contrast, the unsupervised Pointwise model only achieved 89.7% accuracy on the experimental data but reported 78.3% accuracy on the synthetic data. • The best anomaly indicators are higher frequency (400 MHz BW) FDR data. Other tests may be interesting but for this study, this was the best predicter.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND

Improving the National Solar Radiation Database (NSRDB) Using a Physics-Based Direct Normal Irradiance (DNI) Model: Preprint

The National Solar Radiation Database (NSRDB) is a widely used resource providing satellite-derived solar data across the United States and globally. While the NSRDB employs a physical model for computing global horizontal irradiance (GHI), its current method for estimating cloudy-sky direct normal irradiance (DNI) relies on surface observations and empirical models. Recently, a novel physics-based approach, the Fast All-Sky Radiation Model for Solar applications with DNI (FARMS-DNI), was developed to enhance the DNI forecasting. FARMS-DNI incorporates both direct and scattered solar radiation within the circumsolar region, resulting in improved day-ahead DNI predictions when integrated into the Weather Research and Forecasting model with Solar extensions (WRF-Solar). This study integrates FARMS-DNI into the NSRDB algorithm to generate high-resolution DNI data from satellite resources. Our findings reveal that FARMS-DNI effectively mitigates the substantial DNI overestimation present in the conventional NSRDB across surface sites, particularly in conditions categorized as cloudy overcast. Consequently, this innovative model substantially enhances the overall accuracy of the NSRDB.

DNI

Precision Measurement of Neutrino Oscillation Parameters with 10 Years of Data from the NOvA Experiment

This Letter reports measurements of muon-neutrino disappearance and electron-neutrino appearance and the corresponding antineutrino processes between the two NOvA detectors in the NuMI neutrino beam. These measurements use a dataset with double the neutrino mode beam exposure that was previously analyzed, along with improved simulation and analysis techniques. A joint fit to these samples in the three-flavor paradigm results in the most precise single-experiment constraint on the atmospheric neutrino mass-splitting, $Δm^2_{32}= 2.431^{+0.036}_{-0.034} (-2.479^{+0.036}_{-0.036}) \times 10^{-3}$~eV$^2$ if the mass ordering is Normal (Inverted). In both orderings, a region close to maximal mixing with $\sin^2θ_{23}=0.55_{+0.06}^{-0.02}$ is preferred. The NOvA data show a mild preference for the Normal mass ordering with a Bayes factor of 2.4 (corresponding to 70% of the posterior probability), indicating that the Normal ordering is 2.4 times more probable than the Inverted ordering. When incorporating a 2D $Δm^2_{32}\textrm{--}\sin^2 2θ_{13}$ constraint based on Daya Bay data, this preference strengthens to a Bayes factor of 6.6 (87\%).

FOS: Physical sciences

Data from: "Towards CONUS-Wide ML-Augmented Conceptually-Interpretable Modeling of Catchment-Scale Precipitation-Storage-Runoff Dynamics"

This data package was generated to support the manuscript “Towards CONUS-Wide Machine Learning-Augmented Conceptually Interpretable Modeling of Catchment-Scale Precipitation-Storage-Runoff Dynamics.” It provides input files, model outputs, plotting data, scripts, notebooks, and documentation used to develop, evaluate, and reproduce Mass-Conserving Perceptron (MCP)-based hydrologic modeling experiments across 513 selected Catchment Attributes and Meteorology for Large-sample Studies in the United States (CAMELS-US) basins. The files are organized by modeling component and analysis purpose, including rainfall–runoff experiments, snow module experiments, coupled hydrologic-snow experiments, Long Short-Term Memory (LSTM) benchmark results, model skill metrics, initialization and epoch records, cell-state normalization files, Akaike Information Criterion (AIC)-based model comparison files, and data used to generate manuscript figures. Tabular files can be opened using standard spreadsheet software or Python/R data-analysis tools. Python scripts, Jupyter notebooks, and selected MATLAB scripts are included for model execution, postprocessing, plotting, and statistical analysis. Quality assurance and quality control were conducted through the source-data selection and modeling workflow. Meteorological forcing, streamflow, and static catchment attributes were derived from the CAMELS-US dataset, and snow water equivalent data were derived from the University of Arizona (UA) Snow Water Equivalent dataset. Selected basins and time periods were screened during the associated research workflow to avoid missing observations or poor-quality cases. Static geospatial features were processed primarily using Quantum Geographic Information System (QGIS) and Geospatial Data Abstraction Library (GDAL) workflows. Additional details are provided in the associated manuscript and documentation.

ESS-DIVE CSV File Formatting Guidelines Reporting

TransPlatformer

We propose TransPlatformer for translating toxicogenomics from one platform to another. Transcriptomic profiling has evolved through multiple generations of technology, from microarrays (e.g., Affymetrix, CodeLink) to more recent high-throughput sequencing and targeted panels such as S1500+. Microarrays, which dominated gene expression studies in the early 2000s, provided affordable and high-throughput transcript quantification but suffered from cross-hybridization issues and limited dynamic range . RNA-Seq, introduced in the late 2000s, revolutionized transcriptomics by enabling unbiased and comprehensive gene expression analysis, albeit at higher costs and computational demands . Despite advances, many studies rely on historical microarray data, necessitating the translation of legacy data into modern platforms to ensure continuity and comparability. This translation is complicated by factors such as platform-specific probe design, differences in transcript coverage, and batch effects . Existing methods for cross-platform mapping include statistical normalization, machine learning models, and biological anchoring approaches. The ability to translate transcriptomic data between platforms has broad implications, including enhanced meta-analyses, improved toxicological modeling, and better integration of historical datasets with contemporary research. TransPlatformer seeks to contribute to this effort by evaluating translation methodologies and proposing novel strategies to improve cross-platform gene expression harmonization. In this repository there are code examples for TransPlatformer implementation

Cong, Guojing

LinkML: an open data modeling framework

Background Scientific research relies on well-structured, standardized data; however, much of it is stored in formats such as free-text lab notebooks, nonstandardized spreadsheets, or data repositories. This lack of structure challenges interoperability, making data integration, validation, and reuse difficult. Findings LinkML (Linked Data Modeling Language) is an open framework that simplifies the process of authoring, validating, and sharing data. LinkML can describe a range of data structures, from flat, list-based models to complex, interrelated, and normalized models that utilize polymorphism and compound inheritance. It offers an approachable syntax that is not tied to any one technical architecture and can be integrated seamlessly with many existing frameworks. The LinkML syntax provides a standard way to describe schemas, classes, and relationships, allowing modelers to build well-defined, stable, and optionally ontology-aligned data structures. Once defined, LinkML schemas may be imported into other LinkML schemas. These key features make LinkML an accessible platform for interdisciplinary collaboration and a reliable way to define and share data semantics. Conclusions LinkML helps reduce heterogeneity, complexity, and the proliferation of single-use data models while simultaneously enabling compliance with FAIR (Findable, Accessible, Interoperable, and Reusable) data standards. LinkML has seen increasing adoption in various fields, including biology, chemistry, biomedicine, microbiome research, finance, electrical engineering, transportation, and commercial software development. In short, LinkML makes implicit models explicitly computable and allows data to be standardized at their origin. LinkML documentation and code are available at https://linkml.io/.

AI-ready data

Liquid–Liquid Equilibrium Prediction in Fast Pyrolysis Bio-Oil Systems: A Framework for Incorporating Bio-Oil Complexity

The study of mixtures of bio-oil, water and organic solvents in different proportions can serve as a cost-effective analysis of its content due to the formation of immiscible phases. This manuscript attempts to replicate experimentally determined partition coefficients (K OW ) of relevant species present in fast pyrolysis bio-oil (FPBO). A commercial flowsheeting simulator with surrogate bio-oil model representation is used. Concurrently, pyrolytic lignins in FPBO (‘pyrolignin’) do not have an agreed-upon structural representation, and the literature is ripe with wide variations of said representations. Thus, during the description of FPBO, this pyrolignin fraction was modeled using 20 possible structures (phenolic dimers to tetramers), with the goal of determining the structures for which the experimental data are best described. Two cases were considered: Case 1 normalized the reported experimental mass balance, while Case 2 included the unreported fraction in the mass balance to the total pyroligin. Please, add here a comment on the prediction of the Water oil equilibrium. The best KOW predictions for levoglucosan (LVG) were obtained when the system was modeled with no pyrolignin, presenting an MRE under 10% for both systems WO and BO. Among the possible structures, D2 (dimer), F1 (trimer) and I1, and I3 (tetramers) presented MRE ≤ 13% for both cases.

09 BIOMASS FUELS

First Differential Measurement of the Single 𝜋 + Production Cross Section in Neutrino Neutral-Current Scattering

Since its first observation in the 1970s, neutrino-induced neutral-current single positive pion production (NC⁢1⁢𝜋 + ) has remained an elusive and poorly understood interaction channel. This process is a significant background in neutrino oscillation experiments and studying it further is critical for the physics program of next-generation accelerator-based neutrino oscillation experiments. In this Letter, we present the first double-differential cross-section measurement of NC⁢1⁢𝜋 + interactions using data from the ND280 detector of the T2K experiment collected in 𝜈-beam mode. The measured flux-averaged integrated cross section is 𝜎 = (6.07 ± 1.22) × 10 −41 cm 2 /nucleon. We compare the results on a hydrocarbon target to the predictions of several neutrino interaction generators and final-state-interaction models. While model predictions agree with the differential results, the data show a weak preference for a cross-section normalization approximately 30% higher than predicted by most models studied in this Letter.

Neutrino detection

Accurate field-level weak lensing inference for precision cosmology

We present miko, a catalog-to-cosmology pipeline for general flat-sky field-level inference, which provides access to cosmological information beyond the two-point statistics. In the context of weak lensing, we identify several new field-level analysis systematics (such as aliasing, Fourier mode-coupling, and density-induced shape noise), quantify their impact on cosmological constraints, and correct the biases to a percent level. Next, we find that model misspecification can lead to both absolute bias and incorrect uncertainty quantification for the inferred cosmological parameters in realistic simulations. The Gaussian map prior infers unbiased cosmological parameters, regardless of the true data distribution, but it yields overconfident uncertainties. The log-normal map prior quantifies the uncertainties accurately, although it requires careful calibration of the shift parameters for unbiased cosmological parameters. Here, we demonstrate systematics control down to the 2% level for both models, making them suitable for ongoing weak lensing surveys.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS