Engineering PapersSearch

SEARCH · Engineering Papers

Results for “general physics”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 847 records · Page 47

Scaleup and Site-Specific Engineering Design for Global Thermostat Direct Air Capture Technology

The overall goal of the project is completion of an initial design of a commercial-scale, Carbon Capture, Utilization, and Storage Direct Air Capture (CCUS-DAC) plant design at three different sites that captures a net of at least 100,000 tonnes per year (TPY) carbon dioxide (CO 2 ) from the atmosphere and considers compression and conditioning of the captured CO 2 for purpose of pipeline transportation to different geological formation sites for deep well injection and underground storage. In addition to the leading system consisting of a scaled-up Global Thermostat DAC unit, overall plant design includes a combined Heat and Power (CHP) unit integrated with a conventional liquid amine-based carbon capture system (90% capture) and compression facilities. These are considered balance of plant (BOP) systems and are required to provide low carbon-intensity steam and power to the DAC process. A second approach involving provision of DAC units modified to directly capture emissions from the CHP unit was initially assessed as an alternative and it was determined to be less mature than considered approach and not included in the scaled-up plant design efforts. Three geographically diverse continental United States locations were selected to better understand the effect of local/regional ambient conditions on scaled-up DAC system performance and project costs: Bucks, Alabama (hot wet climate), Odessa, Texas (dry hot climate), and Goose Creek, Illinois (mid continental climate). The team focused on initial engineering design activities, including the development of project design criteria, initiation of site-specific studies and investigations, completion of DAC system process and equipment design, and definition of balance of plant (BOP) engineering. The purpose of the activities were to develop Technoeconomic Analysis (TEA), Life-Cycle Analysis (LCA), and Environmental, Health, and Safety (EH&S) analysis, and Business Case Analysis (BCA) to validate that the project engineering and scale-up plans of the DAC systems, at each of the three distinct case studies considered, are technically, economically, and environmentally feasible for commercial-scale operation.

20 FOSSIL-FUELED POWER PLANTS

A Self Consistent 2D Simulation of Coherent Synchrotron Radiation Effects on Beam Dynamics

An increasing interest in high quality and high current electron beams necessitates a thorough understanding and prediction of coherent synchrotron radiation effects. The self-interaction of charged particles in a beam undergoing synchrotron motion is a physically significant process that is all too often computationally intensive with very little analytical results to rely on for the general case. The coherent spectrum of this interaction is of utmost importance to the design of free electron lasers (FELs) and an accurate assessment is imperative for their design. This work presents a novel implementation to the numerical simulation of charged particle beams. The simulation is a self-consistent approach including the self-fields generated by the beam of which coherent synchrotron radiation effects are of primary interest. A particle-in-cell model is used where a planar beam sampled by point particles is deposited on an encompassing grid at each timestep. The electromagnetic fields are calculated on the grid using the retarded potentials according to causality. The electromagnetic forces from the fields are interpolated on each particle which in turn advance in time. The simulation is benchmarked against well-established results for coherent synchrotron radiation effects. In addition, studies are provided that show the convergence of simulation results for increasing resolution. A study into the transverse beam size effects on beam dynamics is performed as well as a proof of concept where the simulation is used by a genetic algorithm to optimize the design parameters of a beam lattice. The results of these studies in tandem verify the efficacy of the simulation for its practical use in accelerator design or the study of synchrotron radiation effects

Duffin, Dallan [Old Dominion Univ., Norfolk, VA (U

Constraints on neutrino physics from DESI DR2 BAO and DR1 full shape

The Dark Energy Spectroscopic Instrument (DESI) Collaboration has obtained robust measurements of baryon acoustic oscillations in the redshift range 0.1 < 𝑧 < 4.2, based on the Lyman-𝛼 forest and galaxies from data release 2. We combine these measurements with cosmic microwave background (CMB) data from Planck and the Atacama Cosmology Telescope to place our tightest constraints yet on the sum of neutrino masses. Assuming the cosmological Λ⁢ CDM model and three degenerate neutrino states, we find ∑𝑚 𝜈 < 0.0642 eV (95%) with a marginalized error of 𝜎⁡(∑𝑚 𝜈 ) = 0.020 eV. We also constrain the effective number of neutrino species, finding 𝑁 eff = 3.2⁢3$^{+0.35}_{−0.34}$ (95%), in line with the Standard Model prediction. When accounting for neutrino oscillation constraints, we find a preference for the normal mass ordering and an upper limit on the lightest neutrino mass of 𝑚 𝑙 < 0.023 eV (95%). However, we determine using frequentist and Bayesian methods that our constraints are in tension with the lower limits derived from neutrino oscillations. Correcting for the physical boundary at zero mass, we report a 95% Feldman-Cousins upper limit of ∑𝑚 𝜈 < 0.053 eV, breaching the lower limit from neutrino oscillations. Considering a more general Bayesian analysis with an effective cosmological neutrino mass parameter, ∑𝑚 𝜈,eff , that allows for negative energy densities and removes unsatisfactory prior weight effects, we derive constraints that are in 3⁢𝜎 tension with the same oscillation limit, while the error rises to 𝜎⁡(∑𝑚 𝜈,eff ) = 0.053 eV. In the absence of unknown systematics, this finding could be interpreted as a hint of new physics not necessarily related to neutrinos. The preference of DESI and CMB data for an evolving dark energy model offers one possible solution. In the 𝑤 0 ⁢𝑤 𝑎 ⁢CDM model, we find ∑𝑚 𝜈 < 0.163 eV (95%), relaxing the neutrino tension. These constraints all rely on the effects of neutrinos on the cosmic expansion history. Using full-shape power spectrum measurements of data release 1 galaxies, we place complementary constraints that rely on neutrino free streaming. Our strongest such limit in Λ ⁢CDM, using selected CMB priors, is ∑𝑚 𝜈 < 0.193 eV (95%).

79 ASTRONOMY AND ASTROPHYSICS

Regulation compliant AI for fusion: explainable image-based feedback control of divertor detachment in DIII-D tokamak

While artificial intelligence (AI) has been promising for fusion control, its inherent black-box nature will make compliant implementation in regulatory environments a challenge. This study implements and validates a real-time AI-enabled linear and interpretable control system for successful divertor detachment control with the DIII-D lower divertor camera. Using D 2 gas, we demonstrate successful feedback divertor detachment control with a mean absolute difference of 2% from the target for both detachment and reattachment. This automatic training and linear processing framework can be extended to any image-based diagnostic for future fusion reactors.

computer vision

A kinetic line-driven radiation operator and its application to Gyrokinetics

A velocity dependent, kinetic model for line radiation is developed for continuum kinetic codes. It has been implemented in the full-f gyrokinetic code Gkeyll. The total radiation for a charge state is modeled as an advection in velocity space with a form of $\nabla_v \cdot(v\nu(v)f(v))$, guaranteeing particle conservation. The velocity dependence (in the form of an effective frequency $\nu(v)$) is found through fitting the energy loss of the operator, i.e. the second velocity moment, to the radiation data in the OpenADAS database. Therefore, each individual transition does not need to be evaluated every time step, significantly reducing the computational cost of including line radiation in a kinetic model. The dependence on velocity instead of the usual, temperature, allows the radiation to be computed from non-Maxwellian electron distribution functions: We benchmark the model against a collisional radiative model using isotropic non-Maxwellian distribution functions. A velocity dependent model of radiation can more accurately describe the radiation in the more kinetic regimes expected in reactor-scale devices. The velocity dependence qualitatively captures the quantum mechanical need for a minimum velocity before any radiation occurs.

kinetic

Towards a NEAMS-based high-fidelity model of the MARVEL reactor

This report outlines the progress of Idaho National Laboratory in developing a high-fidelity and high-resolution model of the Microreactor Applications Research Validation and Evaluation reactor. The model was developed under the Nuclear Energy Advanced Modeling and Simulation microreactor application driver at Idaho National Laboratory. The overarching objective of this activity is the development of a high-fidelity multiphysics MARVEL model using NEAMS tools, and to verify and validate NEAMS tools against MARVEL reference simulation and experimental data, respectively. This is a unique opportunity to conduct multiphysics analysis on a soon-to-be-deployed microreactor. This multiphysics model developed under the NEAMS-funded INL microreactor application driver leverages three single-physics models coupled via the MOOSE’s MultiApp and Transfer systems. The latter systems enable in-memory data transfer between MOOSE-based and MOOSE-wrapped applications. The first single-physics model, that functions as main application, leverages Griffin to model the neutron transport in the core through the discontinuous finite element (DFEM) discrete ordinates solver (SN). Several optimization flags that were developed by the Griffin developer team were beta-tested to enhance the solver’s performance. These include the combined use of using_average_xs and update_averaged_xs_on that enable to avoid expensive on-the-fly cross sections evaluations at each linear iterations in favor of evaluations of the macroscopic cross sections at each Picard iteration. The second single-physics model uses BISON to handle solid heat transfer and asymptotic hydrogen redistribution analysis in the fuel. While the model returns consistent results for the temperature and hydrogen distribution in the fuel, a mismatch was noticed in the calculated temperature in the reflector due to the value of the gap conductance used in our model. Ongoing investigations are being performed to assess the origin of this discrepancy. Finally, the System Analysis Module (SAM) was used to model the flow of the sodium-potassium eutectic in the primary loop. A first verification was also performed showing good agreement in terms of mass flow rate and inlet temperature. All mesh files were generated using the MOOSE Reactor module, removing the need for external meshing tools. Notably, this workscope represents one of the initial applications of the MOOSE Reactor module for modeling highly irregular geometries. The use of the reactor module significantly streamlined the mesh generation process. The full multiphysics mode, that combines all the single physics models, was leveraged to conduct initial steady-state multiphysics simulations to compute power, and temperature distribution in the reactor. Initial testing was performed for transient simulations as well. In this case, the new checkpoint restart capability for eigenvalue calculations was tested showing the capability for streamlined restart of transient calculations. Future work will focus on improving the fidelity of the model by performing comprehensive code-to-code comparisons. For instance, the full-core Griffin neutronics model will be benchmarked against MCNP reference results, that were provided by the MARVEL design team. Additionally, the SAM T/H model will be verified against reference RELAP-5 results for selected accident scenarios. Besides code-to-code verification exercises, the model fidelity will be improved by replacing the single-channel SAM model with a more complex SAM-Pronghorn coupled model, in which the sub-channel capability is deployed to obtain radial temperature resolution in the coolant. This model will be developed in synergy with the NEAMS thermal hydraulics team.

22 GENERAL STUDIES OF NUCLEAR REACTORS

First DIII-D-West hybrid scenario similarity experiments for iter-relevant long-pulse operation

For the first time, similarity experiments between DIII-D and WEST were performed in the ITER "hybrid-like" regime during dedicated campaigns in April and May 2025. The matched parameters include elongation, triangularity, ion ∇B drift direction toward the X-point, qprofile, and core normalized physics quantities in terms of normalized pressure, normalized gyroradius, electron collisionality, ratio of ion to electron temperature, T i /T e . Core transport physics is explored with different aspect ratio (R/a) values (typically 3 at DIII-D and 5 on WEST). DIII-D explored high-beta conditions (electromagnetic effect) with low torque injection (~0 ± 0.5 N•m) using high heating power (up to 6 MW NBI and 2 MW ECRH powers), while scanning the heating mix (ion vs electron), beta, T i /T e , core radiation via controlled tungsten injection using the Laser Blow-Off system. WEST extended operation toward long-duration pulses using its actively cooled tungsten divertor, achieving dominated electron heating regimes with reduced tungsten contamination. Boron impurity injection were scanned on WEST to control edge conditions and core performance. It is found that core confinement improves-manifested by higher electron temperature, total energy content, neutron rate, and ion temperatureunder conditions of low separatrix density, consistent with previous observations [Bourdelle et al., Nucl. Fusion 63 (2023) 056021]. Conditions for Hmode access and for ion heating in electron-dominated regimes in both WEST and DIII-D will be discussed and compared. The ratio of the thermal energy confinement time (τ E ) to the volume-averaged electron-ion collisional heat exchange time (τ e-i ) is a key parameter to enhance ion heating and potentially facilitate H-mode access in electron-heated regimes. These first-of-a-kind coordinated DIII-D and WEST experiments provide a unique multi-machine dataset to validate predictive models and to optimize ITER hybrid-scenario performance under diverse core and edge conditions.

DIII-D

Beyond interpolation: Physics-inspired gating transformers for extrapolating irradiation conditions to novel nuclear fuels

The qualification of advanced nuclear fuels relies on irradiation experiments in test reactors that emulate commercial conditions. Designing these tests requires accurate prediction of key irradiation quantities, particularly heat generation rate and burnup, yet obtaining them typically involves computationally expensive multi-step simulation workflows. We propose a physics-inspired gating transformer (PIGT) that integrates an inverse-square, distance-based attenuation into the encoder representation to bias attention toward physically relevant spatial relationships while retaining data-driven flexibility. Using MiniFuel irradiation data from the High Flux Isotope Reactor at Oak Ridge National Laboratory, we benchmark against ensemble methods, feedforward and recurrent networks, convolutional models, and standard transformers. While baseline models perform well under interpolation, they exhibit a pronounced generalization gap when evaluated on fuels not included in the training set. The proposed model consistently improves extrapolative accuracy and stability, yielding the strongest performance on unseen fuel configurations. These results indicate that a lightweight physics structure embedded within attention mechanisms can substantially improve robustness, enabling more reliable surrogate predictions to accelerate the design of nuclear fuel irradiation experiments.

Fuel qualification

Simulation Center for Runaway Electron Avoidance and Mitigation (SCREAM SciDAC) (Technical Final Report)

Runaway electrons can severely damage the plasma facing components on ITER during a major disruption and pose a major risk for tokamak fusion. It has been recognized that an adequate disruption mitigation system (DMS) is essential for the safe operation of ITER. The United States is responsible for the design and implementation of the disruption mitigation system on ITER, and in July 2016 the Simulation Center for Runaway Electron Avoidance and Mitigation (SCREAM) was launched by DOE, in a joint Fusion Energy Sciences (FES) and Advanced Scientific Computing Research (ASCR) collaboration. SCREAM was a comprehensive theory and simulation SciDAC center that provided physics guidance in the avoidance and mitigation of runaway electrons, and in tandem with domestic and international experiments, helped establish the qualitative and quantitative bases for safe operational scenarios and viable mitigation techniques. The SCREAM center assembled a national team of experts in runaway electron physics, tokamak disruptions, magnetohydrodynamic (MHD) simulation, and advanced algorithms and computing. The team combined advanced simulation and analysis capability facilitated by direct participation of ASCR SciDAC institutes with theoretical models and code development by FES scientists to focus on the runaway risk for ITER and tokamaks in general. The research scope was focussed on integrated simulations of kinetic runaway electrons, including MHD and fluid models of impurity transport, within a research plan guided by theory. The specific research tasks were (1) establish the fundamental physics of runaway generation, saturation, and dynamical evolution in a tokamak; (2) examine the critical path toward runaway avoidance; and (3) investigate the viability and effectiveness of the leading candidate schemes for runaway mitigation. In all three areas, members of the team carried out scoping studies that established the readiness for rapid and critical advances, especially in the deployment and further development of large-to extreme-scale simulation tools. Our multi-pronged computational approach included (1) relativistic Fokker-Planck solvers with discretization in phase space, (2) self-consistent particle-in-cell techniques, (3) particle-based Monte-Carlo, and (4) MHD-particle hybrid simulations. Cross-check between these different methods provided an additional means for verification and further bolstered the fidelity of our physics prediction. Validation against experimental results brings confidence to the predictive capability for ITER and frequently leads to new ideas for understanding and mitigating the thermal quench driven runaway electron phenomenon.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY

Preliminary Activation Wire Analysis for the POKER Experiment Flattop Irradiation

Activation foils, or in this case wires, are commonly deployed to critical experiments at the Na tional Criticality Experiments Research Center (NCERC) to provide additional neutron diagnostics and provide additional energy-dependent responses to tie back to neutronic and/or multi-physics simulations.

07 ISOTOPE AND RADIATION SOURCES

Parametric matrix models

We present a general class of machine learning algorithms called parametric matrix models. In contrast with most existing machine learning models that imitate the biology of neurons, parametric matrix models use matrix equations that emulate physical systems. Similar to how physics problems are usually solved, parametric matrix models learn the governing equations that lead to the desired outputs. Parametric matrix models can be efficiently trained from empirical data, and the equations may use algebraic, differential, or integral relations. While originally designed for scientific computing, we prove that parametric matrix models are universal function approximators that can be applied to general machine learning problems. After introducing the underlying theory, we apply parametric matrix models to a series of different challenges that show their performance for a wide range of problems. For all the challenges tested here, parametric matrix models produce accurate results within an efficient and interpretable computational framework that allows for input feature extrapolation.

Computational science

Machine Learning for Mapping Multipactor Susceptibility in RF Systems: Capabilities and Generalization Constraints

Multipactor is a surface-driven electron avalanche phenomenon that degrades the performance and reliability of radio-frequency (RF) systems in particle accelerator and vacuum electronics applications. Multipactor behavior in a given device structure is conventionally assessed through susceptibility charts, which provide a parameter-space characterization of the instability. In this work, we assess the capabilities of machine-learning (ML) models to learn and predict such susceptibility charts and analyze the constraints governing their generalization across materials. Using a simulation-derived dataset spanning six distinct secondary-electron-yield material profiles in a canonical two-surface planar geometry, we train supervised regression models and artificial neural networks to predict the time-averaged electron growth rate, δavg, across the relevant parameter space. Model performance is evaluated using metrics that explicitly probe the structure of susceptibility charts, including Intersection over Union, Structural Similarity Index, and correlation analysis. Tree-based ensemble models outperform neural-network models in reconstructing susceptibility regions and in generalizing across material domains. Principal-component analysis reveals disjoint material feature distributions, indicating that the piecewise mode structure of multipactor susceptibility is difficult to represent with a single global model and that generalization is constrained by data coverage rather than by model complexity. An exhaustive reduced-coverage study further shows that sparse material-space coverage can yield mean performance in the same general range but producing large variability in the susceptibility-region overlap. These results clarify the capabilities of ML-based surrogate models for parameter-space characterization of multipactor discharge. They also provide guidance for their appropriate use in RF system design.

43 PARTICLE ACCELERATORS

Separation Process of Plant Fibers for Textile and Composite Application: A Review of Recent Advances

Plant fiber resources have gained significant attention for value-added utilization due to their renewability, sustainability, abundance, and widely acknowledged physical properties. The efficient and pragmatic separation of plant fibers is a critical process for their efficient utilization, yet a substantial gap persists between laboratory research advancements and their commercialization. To increase the possibility of research advancements for industrial application, this review summarizes the recent advances in different extraction methodologies of plant fiber research in textile and composite fields. It systematically outlines, compares, and contrasts physical (cryogenic, supercritical carbon dioxide, ultrasonic, steam explosion and microwave heating treatment), chemical (alkali, oxidation, organic solvents and deep eutectic solvents methods), and biological (natural retting, enzymatic and microorganism approaches) methods, addressing their respective mechanisms, strengths, limitations, research progress, and future prospects. In general, traditional chemical approaches have proven significantly effective but are accompanied by high pollution. Conversely, novel chemical treatments such as deep eutectic solvents and organic solvents offer a promising blend of efficiency and environmental friendliness but require deeper studies currently. Meanwhile, physical and biological treatments, though largely eco-friendly, tend to suffer from lower separation efficiencies. The research needs and future direction are also addressed to bridge the gap between scientific advancements and their widespread industrial application.

60 APPLIED LIFE SCIENCES

“Best” Iterative Coupled-Cluster Triples Model? More Evidence for 3CC

To follow up on the unexpectedly good performance of several coupled-cluster models with approximate inclusion of 3-body clusters we performed a more complete assessment of the 3CC method for accurate computational thermochemistry in the standard HEAT framework. New spin-integrated implementation of the 3CC method applicable to closed- and open-shell systems utilizes a new automated toolchain for derivation, optimization, and evaluation of operator algebra in many-body electronic structure. We found that with a double-ζ basis set the 3CC correlation energies and their atomization energy contributions are almost always more accurate (with respect to the CCSDTQ reference) than the CCSDT model as well as the standard CCSD(T) model. The mean absolute errors in cc-pVDZ {3CC, CCSDT, and CCSD(T)} electronic (per valence electron) and atomization energies relative to the CCSDTQ reference for the HEAT data set, were {24, 70, 122} μE h /e and {0.46, 2.00, 2.58} kJ/mol, respectively. The mean absolute errors in the complete-basis-set limit {3CC, CCSDT, and CCSD(T)} atomization energies relative to the HEAT model reference, were {0.52, 2.00, and 1.07} kJ/mol, The significant and systematic reduction of the error by the 3CC method and its lower cost than CCSDT suggests it as a viable candidate for post- CCSD(T) thermochemistry applications, as well as the preferred alternative to CCSDT in general.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Exploring the fusion power plant design space: comparative analysis of positive and negative triangularity tokamaks through optimization

The optimal configuration choice between positive triangularity (PT) and negative triangularity (NT) tokamaks for fusion power plants hinges on navigating different operational constraints rather than achieving specific plasma performance metrics. This study presents a systematic comparison using constrained multi-objective optimization with the integrated FUsion Synthesis Engine (FUSE) framework. Over 200 000 integrated design evaluations were performed exploring the trade-offs between capital cost minimization and operational reliability (maximizing q 95 ) while satisfying engineering constraints including 250 ± 50 MW net electric power, tritium breeding ratio > 1.1, power exhaust limits and an hour flattop time. Both configurations achieve similar cost-performance Pareto fronts through contrasting design philosophies. PT, while demonstrating resilience to pedestal degradation (compensating for up to 40% reduction), are constrained to larger machines (R 0 > 6.5 m) by the narrow operational window between L–H threshold requirements and the research-established power exhaust limit (P sol /R < 15 MW m –1 ). This forces optimization through comparatively reduced magnetic field (∼8 T). NT configurations exploit their freedom from these constraints to access compact, high-field designs (R 0 ~ 5.5 m, B 0 >12 T), creating natural synergy with advancing HTS technology. Sensitivity analyses reveal that PT’s economic viability depends critically on uncertainties in L–H threshold scaling and power handling limits. Notably, a 50% variation in either could eliminate viable designs or enable access to the compact design space. These results suggest configuration selection should be risk-informed: PT offers the lowest-cost path when operational constraints can be confidently predicted, while NT is robust to large variations in constraints and physics uncertainties.

FUSE framework

Optimizing entanglement and Bell inequality violation in top antitop events

A top quark and an antitop quark produced together at colliders have correlated spins. These spins constitute a quantum state that can exhibit entanglement and violate Bell’s inequality. In realistic collider experiments, most analyses allow the axes, as well the Lorentz frame, to vary event by event, thus introducing a dependence on the choice of event-dependent basis leading us to adopt “fictitious states,” rather than genuine quantum states. The basis dependence of fictitious states allows for an optimization procedure, which makes the usage of fictitious states advantageous in measuring entanglement and Bell inequality violation. In this work, we show analytically that the basis that diagonalizes the spin-spin correlations is optimal for maximizing spin correlations, entanglement, and Bell inequality violation. We show that the optimal basis is approximately the same as the fixed beam basis (or the rotated beam basis) near the t t ¯ production threshold, while it approaches the helicity basis far above threshold. Using this basis, we present the sensitivity for entanglement and Bell inequality violation in t t ¯ events at the Large Hadron Collider (LHC) and a future e + e − collider. Since observing Bell inequality violation appears to be quite challenging experimentally, and requires a large dataset in collider experiments, choosing the optimal basis is crucially important to observe Bell inequality violation. Our method and general approach are equally applicable to other systems beyond t t ¯ , including interactions beyond the Standard Model. Published by the American Physical Society 2025

Cheng, Kun (ORCID:0000000249592997)

Prospects for Light Dark Matter Searches at Large-Volume Neutrino Detectors

We propose a new approach to search for light dark matter (DM), with keV–GeV mass, via inelastic nucleus scattering at large-volume neutrino detectors such as Borexino, DUNE, Super-K, Hyper-K, and JUNO. The approach uses inelastic nuclear scattering of cosmic-ray boosted DM, enabling a low-background search for DM in these experiments. Large neutrino detectors, with higher thresholds than dark matter detectors, can be used, since the nuclear deexcitation lines are O ( 10 ) MeV . Using a hadrophilic dark-gauge-boson-portal model as a benchmark, we show that the nuclear inelastic channels generally provide better sensitivity than the elastic scattering for a large region of light DM parameter space. Published by the American Physical Society 2024

Dutta, Bhaskar

Applying Transfer Learning for Street-Scale Nuisance Flood Forecasting in Coastal-Urban Cities

An important challenge with Machine Learning (ML) is its transferability; that is, whether a ML model trained on one set of data can be applied to a second set of data without requiring a full re-training of the model. Transfer Learning (TL) addresses this challenge by transferring knowledge learned in the source domain (the data it was trained on) to the target domain (a second set of data that is statistically different but related, which the model was not trained on). This study investigates the use of TL for street-scale nuisance flood forecasting by exploring whether a ML model trained on data collected for one set of streets can effectively forecast flooding for another set of streets in the same city using TL. The envisioned use case is a city deploying a new flood depth monitoring sensor on a street and using TL to apply a ML model, trained on sensor data from an existing flood depth sensor network, to this new street. Eventually, the new flood depth sensor will have a sufficient dataset for training its own ML model, but TL can be used to fill the gap in time while this new dataset is being generated. This method is explored using a Long Short-Term Memory (LSTM) model trained on data for the flood-prone streets of Norfolk City, Virginia. The data used for training includes environmental time series (rainfall, tide), topographic features (Digital Elevation Model (DEM), Topographic Wetness Index (TWI), Depth To Water (DTW)), and street-scale flood depth time series obtained from a high-fidelity physics-based model, acting as a synthetic street-scale stream depth sensor dataset since actual stream depth sensor data is generally unavailable for most cities. A set of 180 flood-prone streets was used to train a base model, while another set of 180 flood-prone streets was used to re-train that model using different TL strategies. The results show that full-weight re-training proved most effective and minimal re-training of only the output layer was insufficient. The advantage of TL was most pronounced when target data was limited, meaning data collected at the new water depth sensor location included generally less than 18 flood events. As target data increased beyond 18 flood events, the benefit of TL diminished relative to training a ML model directly on the local flood events. These findings can assist cities as they implement street-scale flood sensing systems to create accurate forecasts for new sensing locations that do not yet have sufficient data records to train a local ML model.

Roy, Binata [Univ. of Virginia, Charlottesville, V