Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “generalization”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

Adjoint DSMC for nonlinear spatially-homogeneous Boltzmann equation with a general collision model

We derive an adjoint method for the Direct Simulation Monte Carlo (DSMC) method for the spatially homogeneous Boltzmann equation with a general collision law. This generalizes our previous results in Caflisch et al., which was restricted to the case of Maxwell molecules, for which the collision rate is constant. The main difficulty in generalizing the previous results is that a rejection sampling step is required in the DSMC algorithm in order to handle the variable collision rate. We find a new term corresponding to the so-called score function in the adjoint equation and a new adjoint Jacobian matrix capturing the dependence of the collision parameter on the velocities. The new formula works for a much more general class of collision models.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Out-of-distribution generalization for learning quantum dynamics

Abstract Generalization bounds are a critical tool to assess the training data requirements of Quantum Machine Learning (QML). Recent work has established guarantees for in-distribution generalization of quantum neural networks (QNNs), where training and testing data are drawn from the same data distribution. However, there are currently no results on out-of-distribution generalization in QML, where we require a trained model to perform well even on data drawn from a different distribution to the training distribution. Here, we prove out-of-distribution generalization for the task of learning an unknown unitary. In particular, we show that one can learn the action of a unitary on entangled states having trained only product states. Since product states can be prepared using only single-qubit gates, this advances the prospects of learning quantum dynamics on near term quantum hardware, and further opens up new methods for both the classical and quantum compilation of quantum circuits.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

A simple generalization of the energy gap law for nonradiative processes

For more than 50 years, an elegant energy gap (EG) law developed by Englman and Jortner [Mol. Phys. 18, 145 (1970)] has served as a key theory to understand and model the nearly exponential dependence of nonradiative transition rates on the difference of energy between the initial and final states. This work revisits the theory, clarifies the key assumptions involved in the rate expression, and provides a generalization for the cases where the effects of temperature dependence and low-frequency modes cannot be ignored. For a specific example where the low-frequency vibrational and/or solvation responses can be modeled as an Ohmic spectral density, a simple generalization of the EG law is provided. Test calculations demonstrate that this generalized EG law brings significant improvement over the original EG law. Both the original and generalized EG laws are also compared with the stationary phase approximations developed for electron transfer theory, which suggests the possibility of a simple interpolation formula valid for any value of EG.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Machine Learning for Mapping Multipactor Susceptibility in RF Systems: Capabilities and Generalization Constraints

Multipactor is a surface-driven electron avalanche phenomenon that degrades the performance and reliability of radio-frequency (RF) systems in particle accelerator and vacuum electronics applications. Multipactor behavior in a given device structure is conventionally assessed through susceptibility charts, which provide a parameter-space characterization of the instability. In this work, we assess the capabilities of machine-learning (ML) models to learn and predict such susceptibility charts and analyze the constraints governing their generalization across materials. Using a simulation-derived dataset spanning six distinct secondary-electron-yield material profiles in a canonical two-surface planar geometry, we train supervised regression models and artificial neural networks to predict the time-averaged electron growth rate, δavg, across the relevant parameter space. Model performance is evaluated using metrics that explicitly probe the structure of susceptibility charts, including Intersection over Union, Structural Similarity Index, and correlation analysis. Tree-based ensemble models outperform neural-network models in reconstructing susceptibility regions and in generalizing across material domains. Principal-component analysis reveals disjoint material feature distributions, indicating that the piecewise mode structure of multipactor susceptibility is difficult to represent with a single global model and that generalization is constrained by data coverage rather than by model complexity. An exhaustive reduced-coverage study further shows that sparse material-space coverage can yield mean performance in the same general range but producing large variability in the susceptibility-region overlap. These results clarify the capabilities of ML-based surrogate models for parameter-space characterization of multipactor discharge. They also provide guidance for their appropriate use in RF system design.

43 PARTICLE ACCELERATORS↗

Out-of-Distribution Generalization for Learning Quantum Channels with Low-Energy Coherent States

When experimentally learning the action of a continuous-variable quantum process by probing it with inputs, there will often be some restriction on the input states used. One experimentally simple way to probe a quantum channel is to use low-energy coherent states. Learning a quantum channel in this way presents difficulties, due to the fact that two channels may act similarly on low-energy inputs but very differently for high-energy inputs. They may also act similarly on coherent-state inputs but differently on nonclassical inputs. Extrapolating the behavior of a channel for more general input states from its action on the far more limited set of low-energy coherent states is a case of out-of-distribution generalization. To be sure that such generalization gives meaningful results, one needs to relate error bounds for the training set to bounds that are valid for all inputs. We show that for any pair of channels that act sufficiently similarly on low-energy coherent-state inputs, one can bound how different the input-output relations are for any (high-energy or highly nonclassical) input. This proves that out-of-distribution generalization is always possible for learning quantum channels using low-energy coherent states, as long as enough samples are used.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Emergent anomalies and generalized Luttinger theorems in metals and semimetals

Luttinger's theorem connects a basic microscopic property of a given metallic crystalline material, the number of electrons per unit cell, to the volume, enclosed by its Fermi surface, which defines its low-energy observable properties. Such statements are valuable since, in general, deducing a low-energy description from microscopics, which may perhaps be regarded as the main problem of condensed matter theory, is far from easy. In this paper, we present a unified framework which allows one to discuss Luttinger theorems for ordinary metals as well as closely analogous exact statements for topological (semi)metals, whose low-energy description contains either discrete points or continuous line nodes. This framework is based on the 't Hooft anomaly of the emergent charge conservation symmetry at each point on the Fermi surface, a concept recently proposed by Else et al. Here we find that the Fermi surface codimension p plays a crucial role for the emergent anomaly. For odd p, such as ordinary metals (p = 1) and magnetic Weyl semimetals (p = 3), the emergent symmetry has a generalized chiral anomaly. For even p, such as graphene and nodal line semimetals (both with p = 2), the emergent symmetry has a generalized parity anomaly. When restricted to microscopic symmetries, such as U(1) and lattice symmetries, the emergent anomalies imply (generalized) Luttinger theorems, relating Fermi surface volume to various topological responses. The corresponding topological responses are the charge density for p = 1, Hall conductivity for p = 3, and polarization for p = 2. As a by-product of our results, we clarify exactly what is anomalous about the surface states of nodal line semimetals.

36 MATERIALS SCIENCE↗

INSURE: An Information Theory iNspired diSentanglement and pURification modEl for Domain Generalization

Domain Generalization (DG) aims to learn a generalizable model on the unseen target domain by only training on the multiple observed source domains. Although a variety of DG methods have focused on extracting domain-invariant features, the domain-specific class-relevant features have attracted attention and been argued to benefit generalization to the unseen target domain. To take into account the class-relevant domain-specific information, in this paper we propose an Information theory iNspired diSentanglement and pURification modEl (INSURE) to explicitly disentangle the latent features to obtain sufficient and compact (necessary) class-relevant feature for generalization to the unseen domain. Specifically, we first propose an information theory inspired loss function to ensure the disentangled class-relevant features contain sufficient class label information and the other disentangled auxiliary feature has sufficient domain information. Additionally, we further propose a paired purification loss function to let the auxiliary feature discard all the class-relevant information and thus the class-relevant feature will contain sufficient and compact (necessary) class-relevant information. Moreover, instead of using multiple encoders, we propose to use a learnable binary mask as our disentangler to make the disentanglement more efficient and make the disentangled features complementary to each other. We conduct extensive experiments on five widely used DG benchmark datasets including PACS, VLCS, OfficeHome, TerraIncognita, and DomainNet. The proposed INSURE achieves state-of-the-art performance. We also empirically show that domain-specific class-relevant features are beneficial for domain generalization. The code is available at https://github.com/yuxi120407/INSURE .

97 MATHEMATICS AND COMPUTING↗

Out-of-distribution generalization for learning quantum dynamics

Generalization bounds are a critical tool to assess the training data requirements of Quantum Machine Learning (QML). Recent work has established guarantees for in-distribution generalization of quantum neural networks (QNNs), where training and testing data are drawn from the same data distribution. However, there are currently no results on out-of-distribution generalization in QML, where we require a trained model to perform well even on data drawn from a different distribution to the training distribution. Here, we prove out-of-distribution generalization for the task of learning an unknown unitary. In particular, we show that one can learn the action of a unitary on entangled states having trained only product states. Since product states can be prepared using only single-qubit gates, this advances the prospects of learning quantum dynamics on near term quantum hardware, and further opens up new methods for both the classical and quantum compilation of quantum circuits.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Modified Cascading Generalized Inverse Control Allocation

The current aviation revolution towards electric propulsion aircraft (e.g., electric vertical takeoff-and-landing) brings unique control challenges. These vehicles are typically over-actuated (more effectors than desired control outcomes), may require control strategies for the three phases of flight (hover, transition and cruise), and currently have limited electric power availability. These vehicle challenges bring the need for optimal control allocation to the forefront of research. A leading control allocation algorithm, used in current flight vehicles, is the Cascading Generalized Inverse (CGI). Unfortunately, the Cascading Generalized Inverse algorithm is unable to achieve some desired outcomes, it intermittently provides non-optimal allocations, and it may fail to preserve moment direction near maximal achievable outcomes. In this research, the shortcomings of the Cascading Generalized Inverse algorithm are addressed by augmenting the algorithm with Scalar Difference Quadratic unsaturation identification and location at each iteration. Rigorous theory is shown that the Modified Cascading Generalized Inverse performs better at obtaining optimal allocations for all attainable outcomes. Numerical case studies for over-actuated vehicles demonstrate resolution to the aforementioned deficiencies.

Control Allocation↗

Modified Cascading Generalized Inverse Control Allocation

The current aviation revolution towards electric propulsion aircraft (e.g., electric vertical takeoff-and-landing) brings unique control challenges. These vehicles are typically over-actuated (more effectors than desired control outcomes), may require control strategies for the three phases of flight (hover, transition and cruise), and currently have limited electric power availability. These vehicle challenges bring the need for optimal control allocation to the forefront of research. A leading control allocation algorithm, used in current flight vehicles, is the Cascading Generalized Inverse (CGI). Unfortunately, the Cascading Generalized Inverse algorithm is unable to achieve some desired outcomes, it intermittently provides non-optimal allocations, and it may fail to preserve moment direction near maximal achievable outcomes. In this research, the shortcomings of the Cascading Generalized Inverse algorithm are addressed by augmenting the algorithm with Scalar Difference Quadratic unsaturation identification and location at each iteration. Rigorous theory is shown that the Modified Cascading Generalized Inverse performs better at obtaining optimal allocations for all attainable outcomes. Numerical case studies for over-actuated vehicles demonstrate resolution to the aforementioned deficiencies.

Control Allocation↗

Generalized Bayesian Framework for Evaluation of Integral Benchmark Experiments

A recently published generalized Bayesian optimization framework has provided a way to retract any or all of the three common assumptions underlying the conventional Generalized Linear Least Squares (GLLS) optimization method based on the concepts introduced in reference two. These assumptions are: 1. Perfection: The model used for data evaluation and the prior probability distribution function (PDF) of generalized* data are perfect; 2. Normality: The prior and posterior PDF are normal; and 3. Linearity: The model is linear. In this work we outline how the framework in 1 could be directly adopted for improved evaluation of nuclear criticality integral benchmark experiments (IBEs) by: 1. Removing the first assumption alone by utilizing the concept of imperfections introduced in 1 to enable evaluation in the presence of discrepancies between the data and model or of missing covariance information by a GLLS method that will be seen as a generalization of the conventional GLLS method employed by the TSURFER code, and by 2. Removing the remaining two assumptions by implement- ing a Markov Chain Monte Carlo method for computation of the posterior PDF in the SAMPLER code, where TSURFER and SAMPLER are the uncertainty quantification (UQ) codes for IBEs in the SCALE code system based on the GLLS, and the stochastic method, respectively.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Scattering amplitudes and N -body post-Minkowskian Hamiltonians in general relativity and beyond

We present a general framework for calculating post-Minskowskian, classical, conservative Hamiltonians for N non-spinning bodies in general relativity from relativistic scattering amplitudes. Novel features for N > 2 are described including the subtraction of tree-like iteration contributions and the calculation of non-trivial many-body Fourier transform integrals needed to construct position space potentials. A new approach to calculating these integrals as an expansion in the hierarchical limit is described based on the method of regions. As an explicit example, we present the O(G 2 ) 3-body momentum space potential in general relativity as well as for charged bodies in Einstein-Maxwell. The result is shown to be in perfect agreement with previous post-Newtonian calculations in general relativity up to O(G 2 v 4 ). Furthermore, in appropriate limits the result is shown to agree perfectly with relativistic probe scattering in multi-center extremal black hole backgrounds and with the scattering of slowly-moving extremal black holes in the moduli space approximation.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Generalized entropy of gravitational fluctuations

The corrections to holographic entanglement entropy from bulk quantum fields in a classical gravitational background are now well understood. They lead, in particular, to unitary Page curves for evaporating black holes. However, the correct treatment of quantum fluctuations of the metric, including graviton excitations, is a longstanding problem. We provide a gauge-invariant prescription for the generalized entropy of gravitons in anti-de Sitter space in terms of areas and bulk entanglement entropy, generalizing the quantum extremal surface prescription to accommodate fluctuations in the semiclassical spacetime geometry. This task requires a careful treatment of the area operator on the graviton Hilbert space and the definition of a “quantum extremal gauge” in which the extremal surface is unperturbed. It also requires us to determine the correct vacuum modular Hamiltonian for the graviton field, which we fix by requiring that it doesn’t contain a boundary term in extremal gauge. We check our prescription with an explicit computation of the vacuum-subtracted generalized entropy of states containing a graviton in an AdS-Rindler background. Our results exactly match vacuum-subtracted von Neumann entropies for stress-tensor excited states in holographic conformal field theory with d > 2 dimensions. We also use covariant phase space techniques to give a partial proof of our prescription when the entanglement wedge for the background spacetime has a bifurcate Killing horizon. Along the way, we identify a class of perturbative graviton states that have parametrically larger generalized entropy, in the small G N expansion, than any low-energy excitations of an ordinary quantum field.

1/N expansion↗

Generalized symmetry in dynamical gravity

We explore generalized symmetry in the context of nonlinear dynamical gravity. Our basic strategy is to transcribe known results from Yang-Mills theory directly to gravity via the tetrad formalism, which recasts general relativity as a gauge theory of the local Lorentz group. By analogy, we deduce that gravity exhibits a one-form symmetry implemented by an operator U α labeled by a center element α of the Lorentz group and associated with a certain area measured in Planck units. The corresponding charged line operator W ρ is the holonomy in a spin representation ρ, which is the gravitational analog of a Wilson loop. The topological linking of U α and W ρ has an elegant physical interpretation from classical gravitation: the former materializes an exotic chiral cosmic string defect whose quantized conical deficit angle is measured by the latter. We verify this claim explicitly in an AdS-Schwarzschild black hole background. Notably, our conclusions imply that the standard model exhibits a new symmetry of nature at scales below the lightest neutrino mass. More generally, the absence of global symmetries in quantum gravity suggests that the gravitational one-form symmetry is either gauged or explicitly broken. The latter mandates the existence of fermions. Finally, we comment on generalizations to magnetic higher-form or higher-group gravitational symmetries.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Plasma thermal transport with a generalized 8-moment distribution function

Moment equations that model plasma transport require an ansatz distribution function to close the system of equations. The resulting transport is sensitive to the specific closure used, and several options have been proposed in the literature. Two different 8-moment distribution functions can be generalized to form a single-parameter family of distribution functions. The transport coefficients resulting from this generalized distribution function can be expressed in terms of this free parameter. This provides the flexibility of matching the 8-moment model to some validating result at a given magnetization value, such as Braginskii’s transport, or the more recent results of Davies et al. [Physics of Plasma, 28, 012305 (2021)]. Here, this process can be thought of as solving for the 8-moment distribution function that matches the value of a transport coefficient given by a Chapman-Enskog expansion, while retaining the improved physical properties, such as finite propagation speeds and time dependence which belong to the hyperbolic moment models. Since the presented generalized distribution function only has a single free parameter, only a single transport coefficient can be matched at a time. However, this generalization process may be extended to provide multiple free parameters. The focus of this Brief Communication is on the dramatically improved thermal conductivity of the proposed model compared to the two base moment models.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Generalized spin σ -SCF method

We introduce a generalization of the σ-SCF method to approximate noncollinear spin ground and excited single-reference electronic states by minimizing the Hamiltonian variance. The new method is based on the σ-SCF method, originally proposed by Ye et al. [J. Chem. Phys. 147, 214104 (2017)], and provides a prescription to determine ground and excited noncollinear spin states on an equal footing. Our implementation was carried out utilizing an initial simulated annealing stage followed by a mean-field iterative self-consistent approach to simplify the cumbersome search introduced by generalizing the spin degrees of freedom. The simulated annealing stage ensures a broad exploration of the Hilbert space spanned by the generalized spin single-reference states with random complex element-wise rotations of the generalized density matrix elements in the simulated annealing stage. The mean-field iterative self-consistent stage employs an effective Fockian derived from the variance, which is utilized to converge tightly to the solutions. This process helps us to easily find complex spin structures, avoiding manipulating the initial guess. As proof-of-concept tests, we present results for Hn (n = 3–7) planar rings and polyhedral clusters with geometrical spin frustration. We show that most of these systems have noncollinear spin excited states that can be interpreted in terms of geometric spin frustration. These states are not directly targeted by energy minimization methods, which are meant to converge to the ground state. This stresses the capability of the σ-SCF methodology to find approximate noncollinear spin structures as mean-field excited states.

Chemistry↗

A generalization of the shock invariant relationship

Shock invariant relationship, which was conceived for inert shock waves to derive the 4th power relationship between shock pressure and maximum strain rate, is generalized for reactive shock waves such as Chapman–Jouget detonation and shock-induced vaporization. The generalization, based on the first-order reaction models, is a power function relationship between overall dissipated energy (Δe dis ) and reaction time Δτ such that Δe dis Δτ 1/α = constant, where the power coefficient α is found to be in the range of 2/3–4. Experimental data, though scarce, are consistent with the generalization. Implication of the generalization for inert shocks is also considered and suggests a broad range of the 4th power coefficient including an inequality equation that constrains the shock and particle velocity relationship.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Proper time observables of general gravitational perturbations in laser interferometry-based gravitational wave detectors

We present an explicitly gauge-invariant observable of any general gravitational perturbation, ℎ 𝜇⁢𝜈 [not necessarily due to gravitational waves (GWs)], in a laser interferometry-based GW detector, identifying the signature as the proper time elapsed of the beamsplitter observer, between two events: when a photon passes through the beamsplitter, and when the same photon returns to the beamsplitter after traveling through the interferometer arm and reflecting off the far mirror. Our formalism applies to simple Michelson interferometers and can be generalized to more advanced setups. We demonstrate that the proper time observable for a plane GW is equivalent to the detector strain commonly used by the GW community, though now the common framework can be easily generalized for other types of signals, such as dark matter clumps or spacetime fluctuations from quantum gravity. We provide a simple recipe for computing the proper time observable for a general metric perturbation in linearized gravity and explicitly show that it is invariant under diffeomorphisms of the perturbation, as any physical observable should be.

Dark matter detectors↗