Engineering PapersSearch

SEARCH · Engineering Papers

Results for “Data driven”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

Developing Data-Driven Synthetic Infrastructure Models for Resilience Analysis

Research on infrastructure resilience has produced promising methods to simulate and optimize complex networks to improve performance. However, restrictions on sharing infrastructure models and the steep cost of developing and maintaining infrastructure models presents a roadblock to adoption. To overcome this limitation, this research focuses on methods to create data-driven infrastructure models that will help improve infrastructure resilience and security. The analysis couples incomplete utility data, geospatial data, machine learning, and synthetic network generation methods to rapidly develop and update infrastructure models. The methods are validated using realistic utility models and site-specific data, with a focus on Puerto Rico due to its unique infrastructure challenges and available data. This research highlights promising opportunities for the use of synthetic network generation and machine learning to create infrastructure models when very little data is available. Results demonstrate that hybrid methods, which combine sparse utility data with synthetic models, can enhance model accuracy, and machine learning can predict model attributes using training data from other models. However, the complexity of infrastructure systems means that even minor changes in network connectivity can significantly impact simulation results. Resilience analysis using synthetic infrastructure models shows that while some system behaviors are preserved, the magnitude of disruptions may not be accurately represented, indicating the need for more research and validation before using synthetic models for critical infrastructure investment decisions. The framework outlined in this report represents a significant advance to infrastructure model development and could be applied to additional domains and sites. Future research will continue to streamline and validate methods to help reduce roadblocks to resilience analysis.

24 POWER TRANSMISSION AND DISTRIBUTION

Optimization-Based Data-Driven Approach for Detecting Fault Location in Power Systems

In grids with large penetration of converterinterfaced resources (CIRs), measurements of voltage, current, and line parameters can fluctuate significantly during fault conditions. These fluctuations, combined with complex network topologies and extensive system branching, make accurate fault location challenging. Faults, such as short circuits, can cause prolonged outages with serious socio-economic impacts, highlighting the need for rapid fault identification to minimize downtime. However, current fault detection methods—such as relays and digital fault recorders—often relay information too slowly, impeding swift corrective action. Given the limited availability of high-resolution phasor measurement units, this paper introduces an optimization-based observer to estimate fault locations, grid line parameters, and voltages using local CIR measurements. To preserve the confidentiality of CIRs and enhance estimation accuracy, this study uses a black-box model of CIRs. This bottom-up, event-driven approach can enhances protection and control systems through optimized and real-time fault detection. Simulation results show that the optimization-based data-driven observer can accurately detect fault locations and estimate grid states and parameters, providing valuable insights for utilities and operators in grid applications.

Subedi, Sunil [ORNL] (ORCID:000000034069090X)

Optimizing dynamic wireless charging for electric buses: A data-driven approach to infrastructure planning

The network configuration significantly impacts the performance of dynamic wireless charging (DWC) technology for electric buses. Here, this study presents a novel approach to planning charging infrastructure for public transit using data-driven nonconvex mixed-integer optimization. Integrating DWC and charging station technologies reveals a trade-off between enroute and stationary charging times. Our framework optimizes bus frequency settings and transmitter coil arrangements to minimize operational and infrastructure costs. A case study in Chattanooga, Tennessee, demonstrates the method's effectiveness in mitigating range anxiety and reducing charging expenses. This research implies that integrating DWC technology into public transit systems can enhance the feasibility and cost-effectiveness of electric bus operations, promoting sustainable urban mobility.

33 ADVANCED PROPULSION SYSTEMS

Determining Stellar Elemental Abundances from DESI Spectra with the Data-driven Payne

Abstract Stellar abundances for a large number of stars provide key information for the study of Galactic formation history. Large spectroscopic surveys such as the Dark Energy Spectroscopic Instrument (DESI) and LAMOST take median-to-low-resolution (R≲ 5000) spectra in the full optical wavelength range for millions of stars. However, the line-blending effect in these spectra causes great challenges for elemental abundance determination. Here we employDD-Payne, a data-driven method regularized by differential spectra from stellar physical models, to the DESI early data release spectra for stellar abundance determination. Our implementation delivers 15 labels, including effective temperatureT eff , surface gravity log g , microturbulence velocityv mic , and the abundances for 12 individual elements, namely C, N, O, Mg, Al, Si, Ca, Ti, Cr, Mn, Fe, and Ni. Given a spectral signal-to-noise ratio of 100 per pixel, the internal precisions of the label estimates are about 20 K forT eff , 0.05 dex for log g , and 0.05 dex for most elemental abundances. These results agree with the theoretical limits from the Crámer–Rao bound calculation within a factor of 2. The majority of the accreted halo stars contributed by the Gaia–Enceladus–Sausage are discernible from the disk and in situ halo populations in the resultant [Mg/Fe]–[Fe/H] and [Al/Fe]–[Fe/H] abundance spaces. We also provide distance and orbital parameters for the sample stars, which spread over a distance out to ∼100 kpc. The DESI sample has a significantly higher fraction of distant (or metal-poor) stars than the other existing spectroscopic surveys, making it a powerful data set for studying the Galactic outskirts. The catalog is publicly available.

Astronomy & Astrophysics

Constructing Data-Driven Predictions at the Far Detector for NOvA's Neutrino Oscillation Analysis.

NOvA, is a two-detector, long-baseline neutrino oscillation experiment located at Fermilab, Batavia, IL, USA. It is designed primarily to constrain neutrino oscillation parameters using $\nu_\mu \ (\bar{\nu}_\mu)$ disappearance and $\nu_e \ (\bar{\nu}_e)$ appearance data. The Neutrinos at Main Injector (NuMI) beamline at Fermilab provides a high purity 900 KW intense beam of neutrinos and anti-neutrinos to NOvA. The NOvA Near Detector, located 100m underground and 1km away from the beam source, observes the un-oscillated $\nu_\mu \ (\bar{\nu}_\mu)$ and beam $\nu_e \ (\bar{\nu}_e)$ event spectrum. The Far Detector, located in Ash River, MN, USA, is 809 km from the ND and records the oscillated $\nu_e \ (\bar{\nu}_e)$ and the un-oscillated $\nu_\mu \ (\bar{\nu}_\mu)$ event spectrum. NOvA uses a data-driven technique called extrapolation to predict the expected number of $\nu_\mu \ (\bar{\nu}_\mu)$ and $\nu_e \ (\bar{\nu}_e)$ events at the Far Detector using the Near Detector data. The use of data from a functionally equivalent Near Detector provides a powerful constraint on the systematic uncertainties in NOvA neutrino oscillation analyses. As NOvA continues to add data statistics, a robust constraint on systematics becomes more crucial for neutrino oscillation analysis. The details of the NOvA neutrino oscillation analysis framework and how it constrains dominant systematic uncertainties using the Near Detector data will be discussed in this poster.

43 PARTICLE ACCELERATORS

Hybrid Data‐Driven Discovery of High‐Performance Silver Selenide‐Based Thermoelectric Composites

Optimizing material compositions often enhances thermoelectric performances. However, the large selection of possible base elements and dopants results in a vast composition design space that is too large to systematically search using solely domain knowledge. To address this challenge, a hybrid data-driven strategy that integrates Bayesian optimization (BO) and Gaussian process regression (GPR) is proposed to optimize the composition of five elements (Ag, Se, S, Cu, and Te) in AgSe-based thermoelectric materials. Data is collected from the literature to provide prior knowledge for the initial GPR model, which is updated by actively collected experimental data during the iteration between BO and experiments. Within seven iterations, the optimized AgSe-based materials prepared using a simple high-throughput ink mixing and blade coating method deliver a high power factor of 2100 µW m −1 K −2 , which is a 75% improvement from the baseline composite (nominal composition of Ag 2 Se 1 ). In conclusion, the success of this study provides opportunities to generalize the demonstrated active machine learning technique to accelerate the development and optimization of a wide range of material systems with reduced experimental trials.

36 MATERIALS SCIENCE

Data Driven Correlated Noise Simulation for the ICEBERG LArTPC

Accurate electronic-noise simulation is essential for low-energy physics in liquid-argon TPCs. More realistic noise modeling allows us to better tune reconstruction algorithms and more reliably assess and optimize signal-detection thresholds. We present a data-driven noise simulation framework developed for the ICEBERG test stand for DUNE that generates synthetic noise waveforms that reproduce both (i) the measured per-channel magnitude of the Fast Fourier Transform (FFT) and (ii) frequency-dependent channel-to-channel correlations observed in ICEBERG noise data. Using a dedicated noise-only dataset, we build a compact noise model containing per-channel FFT-magnitude targets together with a small set of band-wise cross-wire color matrices. White noise is generated in the frequency domain by drawing circular-symmetric complex Gaussian coefficients with random phases and scaling them to match the measured FFT-magnitude targets, and cross-wire correlations are subsequently imposed using the stored color matrices. The model and algorithm were integrated into the LArSoft + Wire-Cell Toolkit simulation chain and validated by comparing waveform structure, frequency-domain spectra, and band-limited correlation matrices from simulated noise and ICEBERG data. This approach can be extended to other LArTPC operating conditions.

Ghosh, Avik [Iowa State U.]

Transmission Data-Driven User-Defined Model for Inverter-based and Conventional Power Plants

Recent events in Odessa [1], [2] have shed light on the complexities of integrating large Inverter-Based Resource (IBR) plants with the transmission system, prompting NERC to stress continuous performance monitoring by transmission operators. Challenges such as plant control updates, IBR model revisions, Phase-locked loop loss of synchronism, and protection events have been identified, underscoring the need for enhanced monitoring protocols by regulatory bodies. The recent FERC 901 order underscores the importance of accurate data exchange regarding IBRs for reliability studies. However, limited access to IBR plant-related data hampers effective decision-making for transmission operators (TOP). This paper proposes a method for constructing data-driven User-Defined dynamic Models (UDM) for power plants for validating multiple-event data using field measurements from interconnection bus locations. The problem is formulated as a power plant model identification problem and a multi-task learning approach under partial input observability assumptions is proposed in this work. This approach aims to predict aggregated responses of conventional and IBR power plants during various dynamic physical events which is useful for planning studies under diverse disturbance conditions. Ultimately, this methodology emphasizes the importance of plant visibility to operators in addressing power system challenges, facilitating improved planning and operational studies.

Mahapatra, Kaveri [BATTELLE (PACIFIC NW LAB)]

Data-Driven Compositional Optimization in Misspecified Regimes

With a manifold growth in the scale and intricacy of systems, the challenges of parametric misspecification become pronounced. These concerns are further exacerbated in compositional settings, which emerge in problems complicated by modeling risk and robustness. In “Data-Driven Compositional Optimization in Misspecified Regimes,” the authors consider the resolution of compositional stochastic optimization problems, plagued by parametric misspecification. In considering settings where such misspecification may be resolved via a parallel learning process, the authors develop schemes that can contend with diverse forms of risk, dynamics, and nonconvexity. They provide asymptotic and rate guarantees for unaccelerated and accelerated schemes for convex, strongly convex, and nonconvex problems in a two-level regime with extensions to the multilevel setting. Surprisingly, the nonasymptotic rate guarantees show no degradation from the rate statements obtained in a correctly specified regime and the schemes achieve optimal (or near-optimal) sample complexities for general T-level strongly convex and nonconvex compositional problems.

Business & Economics

Data-Driven Analysis of Multipactor Dynamics via Dynamic Mode Decomposition

Multipactor effect is a performance-limiting kinetic plasma effect that can occur in high-power microwave and radio frequency (RF) devices. Multipactor effect is of special concern in vacuum or near-vacuum conditions such as those in particle accelerators and spaceborne devices. In this work, we present a data-driven reduced-order model (ROM) based on dynamic mode decomposition (DMD) for modeling of multipactor effects. We study multipactor effects and the resulting nonlinear harmonic generation by processing high-fidelity data generated from electromagnetic particle-in-cell (EMPIC) simulations using the DMD algorithm. We also investigate time-delay embedding extensions of DMD with improved generalizability and accuracy for modeling the electron plasma current density behavior. Here, the results show that DMD provides valuable insights into multipactor phenomena by extracting relevant modal spatiotemporal patterns and frequencies. In addition, DMD offers the potential to time extrapolate EMPIC simulations at a minimal cost, thereby reducing overall simulation time.

43 PARTICLE ACCELERATORS

Fast data-driven spectrometer with direct measurement of time and frequency for multiple single photons

We present a single-photon-sensitive spectrometer based on a linear array of 512 single-photon avalanche diode detectors with 0.04 nm spectral and 40 ps temporal resolutions. We employ a fast data-driven operation that allows direct measurement of time and frequency for simultaneous single photons, time- and frequency-stamping each single-photon detection. Our results combine excellent temporal and spectral resolution. This work opens numerous applications in quantum photonics, especially when both spectral and temporal properties of single photons can be exploited.

79 ASTRONOMY AND ASTROPHYSICS

Review of data-driven models for quantifying load shed by non-residential buildings in the United States

Shifting and shedding power demand in buildings can be cost-effective techniques for grids to function reliably and for end users to earn compensation. Grid operators reimburse customers in proportion to the quantity of load shed. Simple data-driven methods are used to quantify this shed, which is the difference between a measured load during the event and modeled "baseline" that would have occurred in absence of the event. These methods have evolved over the years and in many cases have been integrated with building physics, to make them a hybrid between physics based and empirical models. However, there is no comprehensive analysis that provides guidance to building operators, grid operators and researchers in selecting appropriate models based on their specific needs and available data. Here, this work aims to fill this gap by critically assessing the performance of baseline models put forward from the year 2000 through 2023. The literature reviewed includes reports generated by grid operators, reports from national laboratories and academic journal articles. The work outlines modeling features like the inputs, training period, estimation method, adjustments to fine tune the predictions and metrics to evaluate the performance. A comprehensive list of 50 models has been provided. For each model, the study explores the applicability of the model to weather sensitive buildings, variability in the building profile, timing of the event, and whether the building reduces energy consumption before an event. The work identifies the situations in which a particular model works and draws lessons based on evidence of performance. Finally, recommendations to aid in model selection are given.

97 MATHEMATICS AND COMPUTING

Absorption dissymmetry factor enhancement: A data-driven approach to unravel the synthesis knobs of chiral 2D perovskites

Chiral 2D metal halide perovskites (MHPs) are promising for spin-optoelectronic applications, yet their absorption dissymmetry factor (g abs ) exhibits significant variability due to complex, co-dependent structural and experimental factors. Here, we established a data-driven framework using Pearson’s correlation, ANOVA, and Gaussian process regression to identify and model key synthesis “knobs” governing these properties. The analysis revealed that solvent choice is the primary factor driving variability. For acetonitrile-based films, g abs was maximized by optimizing annealing temperature and film thickness. Conversely, films from higher boiling point solvents showed complex dependencies on annealing temperature, excitonic integral intensity, and film texture. These statistical correlations provide a roadmap for the rational design of high-performance chiral MHPs and establish a foundation for future machine learning-driven material exploration.

ANOVA

Data-driven prediction of scaling and ignition of inertial confinement fusion experiments

Recent advances in inertial confinement fusion (ICF) at the National Ignition Facility (NIF), including ignition and energy gain, are enabled by a close coupling between experiments and high-fidelity simulations. Neither simulations nor experiments can fully constrain the behavior of ICF implosions on their own, meaning pre- and postshot simulation studies must incorporate experimental data to be reliable. Linking past data with simulations to make predictions for upcoming designs and quantifying the uncertainty in those predictions has been an ongoing challenge in ICF research. We have developed a data-driven approach to prediction and uncertainty quantification that combines large ensembles of simulations with Bayesian inference and deep learning. The approach builds a predictive model for the statistical distribution of key performance parameters, which is jointly informed by past experiments and physics simulations. The prediction distribution captures the impact of experimental uncertainty, expert priors, design changes, and shot-to-shot variations. We have used this new capability to predict a 10× increase in ignition probability between Hybrid-E shots driven with 2.05 MJ compared to 1.9 MJ, and validated our predictions against subsequent experiments. We describe our new Bayesian postshot and prediction capabilities, discuss their application to NIF ignition and validate the results, and finally investigate the impact of data sparsity on our prediction results.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY

Data-Driven Discovery and Experimental Validation of Solvent Polarity Effects on Conjugated Polymer Solution-to-Film Assembly Pathways

Understanding how solvent properties influence the solution-to-film assembly of conjugated polymers remains a critical challenge due to the complex and intertwined nature of polymer–solvent interactions. In this study, we integrate a data-driven framework with experimental validation to identify key parameters influencing the assembly and performance of poly[2,5-(2-octyldodecyl)-3,6-diketopyrrolopyrrole-alt-5,5-(2,5-di(thien-2-yl)thieno[3,2-b]thiophene)] (DPP-DTT) in organic field-effect transistors (OFETs). A machine learning (ML) approach identified the normalized Reichardt polarity parameter (E T N ) as a significant descriptor correlated with DPP-DTT hole mobility (μ). Systematic DPP-DTT devices fabricated using solvents across a wide E T N range revealed that higher E T N solvents yield enhanced μ. To elucidate the structural origins of high μ, we conducted comprehensive analyses using UV–vis–NIR spectroscopy and grazing incidence wide angle X-ray scattering (GIWAXS) measurements. The results revealed that films processed from high E T N solvents exhibit reduced paracrystallinity. By analyzing the solution-state behavior using optical microscopy and solution WAXS, we revealed polymer solubility differences in the various solvents and associated distinct polymer assembly pathways, elucidating why the high E T N solvent produces long-range ordered films. Notably, the high E T N solvent shows a pronounced preference for liquid-crystal (LC)-mediated assembly, providing a mechanistic explanation for the enhanced structural order. Therefore, these results demonstrate that solvent polarity, as evaluated by E T N , serves as an important parameter that plays a significant role in the DPP-DTT assembly pathway and resultant solid-state morphology. This work provides a strategy for integrating data science with experiments to identify critical parameters associated with complex polymer systems and helps guide rational process design for high-performance organic electronics.

36 MATERIALS SCIENCE