Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Vectorization”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 253 records · Page 14

Tagging efficiency study of incoherent diffractive vector meson production at the second interaction region at the Electron-Ion Collider

The Electron-Ion Collider (EIC) is an upcoming accelerator facility aimed at exploring the properties of quarks and gluons in nucleons and nuclei, shedding light on their structure and dynamics. The inaugural experimental apparatus, ePIC (electron-Proton and Ion Collider), is designed as a general-purpose detector to address the National Academy of Sciences and the Nuclear Science Advisory Committee physics program at the EIC. The wider EIC community is strongly supporting a second interaction region and an associated second detector to enhance the full science program. In this study, we evaluate how this second interaction region and detector can be complementary to ePIC. The layout of an interaction region for the second detector offers a secondary focus that provides better forward detector acceptance at scattering angles near θ ~ 0 mrad, which can specifically enhance the exclusive, tagging, and diffractive physics program. Here, this article presents an analysis of a tagging program using the second interaction region layout with incoherent diffractive vector meson production. The current design of the second EIC interaction region is evaluated for its vetoing capabilities of incoherent events required for the study of coherent diffractive measurements. We find an increased vetoing performance compared to the ePIC interaction region, thus improving measurements which are important for the spatial imaging of nucleons and nuclei.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Probing Heavy Majorana Neutrinos and the Weinberg Operator through Vector Boson Fusion Processes in Proton-Proton Collisions at s = 13 TeV

The first search exploiting the vector boson fusion process to probe heavy Majorana neutrinos and the Weinberg operator at the LHC is presented. The search is performed in the same-sign dimuon final state using a proton-proton collision dataset recorded at s = 13 TeV , collected with the CMS detector and corresponding to a total integrated luminosity of 138 fb -1 . The results are found to agree with the predictions of the standard model. For heavy Majorana neutrinos, constraints on the squared mixing element between the muon and the heavy neutrino are derived in the heavy neutrino mass range 50 GeV–25 TeV; for masses above 650 GeV these are the most stringent constraints from searches at the LHC to date. A first test of the Weinberg operator at colliders provides an observed upper limit at 95% confidence level on the effective $μμ$ Majorana neutrino mass of 10.8 GeV.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Light and Strange Vector Resonances from Lattice QCD at Physical Quark Masses

We present the first ab initio calculation at physical quark masses of scattering amplitudes describing the lightest pseudoscalar mesons interacting via the strong force in the vector channel. Using lattice quantum chromodynamics, we postdict the defining parameters for two short-lived resonances, the ρ(770) and K*(892) which manifest as complex energy poles in ππ and Kπ scattering amplitudes, respectively. The calculation proceeds by first computing the finite-volume energy spectrum of the two-hadron systems and then determining the amplitudes from the energies using the Lüscher formalism. The error budget includes a data-driven systematic error, obtained by scanning possible fit ranges and fit models to extract the spectrum from Euclidean correlators, as well as the scattering amplitudes from the latter. The final results, obtained by analytically continuing multiple parametrizations into the complex energy plane, are M ρ = 796(5)(50)MeV, Γ ρ = 192(10)(31) MeV, M K* = 893(2)(54) MeV, and Γ K = 51(2)(11) MeV, where the subscript indicates the resonance and M and Γ stand for the mass and width, respectively, and where the first bracket indicates the statistical and the second bracket the systematic uncertainty.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Observation of an Axial-Vector State in the Study of the Decay ψ ( 3686 ) → ϕ η η ′

Using ( 2712.4 ± 14.3 ) × 10 6 ψ ( 3686 ) events collected with the BESIII detector at BEPCII, a partial wave analysis of the decay ψ ( 3686 ) → ϕ η η ′ is performed with the covariant tensor approach. In addition to the established states h 1 ( 1900 ) and ϕ ( 2170 ) , an axial-vector state with a mass near 2.3 GeV / c 2 is observed for the first time. Its mass and width are measured to be 2316 ± 9 stat ± 3 0 syst MeV / c 2 and 89 ± 1 5 stat ± 2 6 syst MeV , respectively. The product branching fractions of B [ ψ ( 3686 ) → X ( 2300 ) η ′ ] B [ X ( 2300 ) → ϕ η ] and B [ ψ ( 3686 ) → X ( 2300 ) η ] B [ X ( 2300 ) → ϕ η ′ ] are determined to be ( 4.8 ± 1.3 stat ± 0.7 syst ) × 10 − 6 and ( 2.2 ± 0.7 stat ± 0.7 syst ) × 10 − 6 , respectively. The branching fraction B [ ψ ( 3686 ) → ϕ η η ′ ] is measured for the first time to be ( 3.14 ± 0.1 7 stat ± 0.2 4 syst ) × 10 − 5 . The first uncertainties are statistical and the second are systematic. Published by the American Physical Society 2025

Ablikim, M.↗

Search for a Neutral Gauge Boson with Nonuniversal Fermion Couplings in Vector Boson Fusion Processes in Proton-Proton Collisions at $\sqrt{𝑠}$ = 13 TeV

The first search for a heavy neutral spin-1 gauge boson (𝑍′) with nonuniversal fermion couplings produced via vector boson fusion processes and decaying to tau leptons or 𝑊 bosons is presented. The analysis is performed using LHC data at $\sqrt{𝑠}$ = 13 TeV, collected from 2016 to 2018 with the CMS experiment and corresponding to an integrated luminosity of 138 fb −1 . The data are consistent with the standard model predictions. Upper limits are set on the product of the cross section for production of the 𝑍′ boson and its branching fraction to 𝜏⁢𝜏 or 𝑊⁢𝑊. The presence of a 𝑍′ boson decaying to 𝜏 + ⁢𝜏 − (𝑊 + ⁢𝑊 − ) is excluded for masses up to 2.45(1.60) TeV, depending on the 𝑍′ boson coupling to standard model weak bosons, and assuming a 𝑍′→𝜏 + ⁢𝜏 − (𝑊 + ⁢𝑊 − ) branching fraction of 50%.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Coherent Spin Pumping Originated from Sub-Terahertz Néel Vector Dynamics in Easy Plane α -Fe 2 O 3 /Pt

We present a thorough study of spin-to-charge current interconversion in bulk and thin films of (0001) 𝛼-Fe2O3/Pt heterostructures by all-optical polarization-controlled microwave excitation at sub-Terahertz frequencies. Our results demonstrate that coherent spin pumping is generated through excitations of both the q-FM and q-AFM modes of antiferromagnetic resonance, provided that the corresponding selection rules are met. Our findings significantly advance current understanding of spin pumping in easy-plane antiferromagnets not only by claiming a substantial contribution from the Néel-vector dynamics of the q-AFM mode but also by experimentally determining the relative strength of the intra- and cross-sublattice components of the spin-mixing conductance, challenging recent reports where the absence of spin pumping from the q-AFM mode in Hematite was interpreted as a complete cancellation between these two components. We also provide an explanation for the previously reported observations and show how the q-AFM spin pumping actually vanishes for thin films, which we speculate being either due to an increased level of inhomogeneities or to insufficient film thickness for the q-AFM mode to fully realize.

Fritjofson, Gregory [University of Central Florida↗

Efficient Hierarchical State Vector Simulation of Quantum Circuits via Acyclic Graph Partitioning

Early but promising results in quantum computing have been enabled by the concurrent development of quantum algorithms, devices, and materials. Classical simulation of quantum programs has enabled the design and analysis of algorithms and implementation strategies targeting current and anticipated quantum device architectures. In this paper, we present a graph-based approach to achieve efficient quantum circuit simulation. Our approach involves partitioning the graph representation of a given quantum circuit into sub-graphs/circuits that exhibit better data locality. Simulation of each sub-circuit is organized hierarchically, with the iterative construction and simulation of smaller state vectors, improving overall performance. Also, this partitioning reduces the number of passes through data, improving the total computation time. We present three partitioning strategies and observe that acyclic graph partitioning typically results in the best time-to-solution. In contrast, other strategies reduce the partitioning time at the expense of potentially increased simulation times. Experimental evaluation demonstrates the effectiveness of our approach.

Fang, Bo↗

Fast Parallel Tensor Times Same Vector for Hypergraphs

Hypergraphs are a popular paradigm to rep- resent complex real-world networks exhibiting multi-way relationships of varying sizes. Mining centrality in hyper- graphs via symmetric adjacency tensors has only recently become computationally feasible for large and complex datasets. To enable scalable computation of these and related hypergraph analytics, here we focus on the Sparse Symmetric Tensor Times Same Vector (S3TTVC) oper- ation. We introduce the Compound Compressed Sparse Symmetric (CCSS) format, an extension of the compact CSS format for hypergraphs of varying hyperedge sizes and present a shared-memory parallel algorithm to compute S3TTVC. We experimentally show S3TTVC computation using the CCSS format achieves better performance than the naive baseline, and is subsequently more performant for hypergraph H-eigenvector centrality.

Shivakumar, Shruti↗

On the Investigation of Phase Fault Classification in Power Grid Signals: A Case Study for Support Vector Machines, Decision Tree and Random Forest

In monitoring the power grid, an ability to differentiate between fault types is essential to ensuring electrical safety. Accordingly, this study introduces a fault detection and classification method by considering different machine learning (ML) and feature extraction (FE) methods combinations. Specifically, the proposed method is established in two classification layers; the first layer determines the fault, and the second layer distinguishes the type of fault. Based on the proposed system model, this study seeks to determine the influential data attributes in a power grid signal using FE methods, including fast Fourier transform, power spectral density (PSD), auto-correlation, and wavelet transform (WT). A cross-comparison of the effectiveness of the Support Vector Machine (SVM), Decision Tree (DT), and Random Forest (RF) is also performed to accomplish the classification layers of the proposed method. The designed algorithm is analyzed under the various combinations of FE and ML methods, and outcomes are presented by considering the trade-off between computational complexity and prediction accuracy. The results reveal that the RF-based ML algorithm shows the most accurate classification performance with PSD, and the most time-saving of the models is the DT WT. Also, SVM emerges superior on a subsequent test of the simulated models on real-world signals.

Galbraith, Kelli↗

Implementation of a Doppler-Free Saturation Spectroscopy (DFSS) Diagnostic for Helicon Wave Electric Field Vector Measurement in Edge Plasma in DIII-D

A laser-based technique known as Doppler-free saturation spectroscopy (DFSS) has been designed, fabricated, and installed on the DIII-D National Fusion Facility to measure the helicon wave electric field vector in the edge plasma. These experimental measurements quantify phenomena resulting in decreased current drive efficiency due to wave/edge plasma interactions. This implementation of DFSS on DIII-D is the first of its kind on a tokamak and thus presents unique engineering challenges, including integration of the system onto an existing multidiagnostic port flange without impacts to system serviceability, as well as maintaining precise laser alignment over a 2-m distance during disruptions and thermal drift of the vessel. Further, these challenges were resolved using innovative design approaches such as a novel decoupled shutter system to facilitate serviceability of the in-vessel mirror assemblies without the need for personnel vessel entry, as well as an ex-vessel piezo-mirror-based optical system for laser beam shaping and real-time steering of the measurement location. The solutions to these engineering challenges were demonstrated during the successful installation and operation of these diagnostic components during the 2022 DIII-D vent and subsequent experimental campaign.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Situating the Vector Density Approach Among Contemporary Continuum Theories of Dislocation Dynamics

Abstract For the past century, dislocations have been understood to be the carriers of plastic deformation in crystalline solids. However, their collective behavior is still poorly understood. Progress in understanding the collective behavior of dislocations has primarily come in one of two modes: the simulation of systems of interacting discrete dislocations and the treatment of density measures of varying complexity that are considered as continuum fields. A summary of contemporary models of continuum dislocation dynamics is presented. These include, in order of complexity, the two-dimensional statistical theory of dislocations, the field dislocation mechanics treating the total Kröner–Nye tensor, vector density approaches that treat geometrically necessary dislocations on each slip system of a crystal, and high-order theories that examine the effect of dislocation curvature and distribution over orientation. Each of theories contain common themes, including statistical closure of the kinetic dislocation transport equations and treatment of dislocation reactions such as junction formation. An emphasis is placed on how these common themes rely on closure relations obtained by analysis of discrete dislocation dynamics experiments. The outlook of these various continuum theories of dislocation motion is then discussed.

Engineering↗

Butterfly Factorization Via Randomized Matrix-Vector Multiplications

This paper presents an adaptive randomized algorithm for computing the butterfly factorization of an m × n matrix with m ≈ n provided that both the matrix and its transpose can be rapidly applied to arbitrary vectors. The resulting factorization is composed of O(log n) sparse factors, each containing O(n) nonzero entries. The factorization can be attained using O(n 3/2 log n) computation and O(n log n) memory resources. Furthermore, the proposed algorithm can be implemented in parallel and can apply to matrices with strong or weak admissibility conditions arising from surface integral equation solvers as well as multi-frontal-based finite-difference, finite-element, or finite-volume solvers. A distributed-memory parallel implementation of the algorithm demonstrates excellent scaling behavior.

97 MATHEMATICS AND COMPUTING↗

Measurements of Higgs bosons decaying to bottom quarks from vector boson fusion production with the ATLAS experiment at $\sqrt{s}=13\,\text {TeV}$

The paper presents a measurement of the Standard Model Higgs Boson decaying to b-quark pairs in the vector boson fusion (VBF) production mode. A sample corresponding to 126 fb –1 of √s = 13TeV proton–proton collision data, collected with the ATLAS experiment at the Large Hadron Collider, is analyzed utilizing an adversarial neural network for event classification. The signal strength, defined as the ratio of the measured signal yield to that predicted by the Standard Model for VBF Higgs production, is measured to be $0.95$ $^{+0.38}_{–0.36}$, corresponding to an observed (expected) significance of 2.6 (2.8) standard deviations from the background only hypothesis. The results are additionally combined with an analysis of Higgs bosons decaying to b-quarks, produced via VBF in association with a photon.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Searches for exclusive Higgs and Z boson decays into a vector quarkonium state and a photon using 139 fb$^{-1}$ of ATLAS $\sqrt{s}=13$ TeV proton–proton collision data

Searches for the exclusive decays of Higgs and Z bosons into a vector quarkonium state and a photon are performed in the μ + μ - γ final state with a proton– proton collision data sample corresponding to an integrated luminosity of 139 fb -1 collected at $\sqrt{s}=13$ TeV with the ATLAS detector at the CERN Large Hadron Collider. The observed data are compatible with the expected back grounds. The 95% confidence-level upper limits on the branching fractions of the Higgs boson decays into J/ψ γ , ψ(2S) γ , and $Υ$(1S, 2S, 3S) γ are found to be 2.0 × 10 -4 , 10.5×10 -4 , and (2.5, 4.2, 3.4)×10 -4 , respectively, assuming Standard Model production of the Higgs boson. The corresponding 95% CL upper limits on the branching fractions of the Z boson decays are 1.2 × 10 -6 , 2.4 × 10 -6 , and (1.1, 1.3, 2.4) × 10 -6 . An observed 95% CL interval of (-133, 175) is obtained for the κ c /κ γ ratio of Higgs boson coupling modifiers, and a 95% CL interval of (-37, 40) is obtained for κ b /κ γ .

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Search for light long-lived neutral particles from Higgs boson decays via vector-boson-fusion production from pp collisions at $\sqrt{s}=13$ TeV with the ATLAS detector

A search is reported for long-lived dark photons with masses between 0.1 GeV and 15 GeV, from exotic decays of Higgs bosons produced via vector-boson-fusion. Events that contain displaced collimated Standard Model fermions reconstructed in the calorimeter or muon spectrometer are probed. This search uses the full LHC Run 2 (2015–2018) data sample collected in proton–proton collisions at $\sqrt{s}=13$ TeV, corresponding to an integrated luminosity of 139 fb –1 . Dominant backgrounds from Standard Model processes and non-collision sources are estimated using data-driven techniques. The observed event yields in the signal regions are consistent with the expected background. Upper limits on the Higgs boson to dark photon branching fraction are reported as a function of the dark photon mean proper decay length or of the dark photon mass and the coupling between the Standard Model and the potential dark sector. This search is combined with previous ATLAS searches obtained in the gluon–gluon fusion and WH production modes. A branching fraction above 10% is excluded at 95% CL for a 125 GeV Higgs boson decaying into two dark photons for dark photon mean proper decay lengths between 173 and 1296 mm and mass of 10 GeV.

Aad, G. (ORCID:0000000266654934)↗

Combined effective field theory interpretation of Higgs boson, electroweak vector boson, top quark, and multijet measurements

Constraints on Wilson coefficients (WCs) corresponding to dimension-6 operators of the standard model effective field theory (SMEFT) are determined from a simultaneous fit to seven sets of CMS measurements probing Higgs boson, electroweak vector boson, top quark, and multijet production. Measurements of electroweak precision observables are also included and provide complementary constraints to those from the CMS experiment. The CMS measurements, using LHC proton-proton collision data at $\sqrt{s}=13\,\text {Te}\text {V} $, corresponding to integrated luminosities of 36.3 or 138$\,\text {fb}^{-1}$, are chosen to provide sensitivity to a broad set of operators, for which consistent SMEFT predictions can be derived. These are primarily measurements of differential cross sections which are parameterized as functions of the WCs. In measurements targeting ${\text {t}} (\bar{\textrm{t}})\text {X} $ production, SMEFT effects are modelled at the detector level. Individual constraints on 64 WCs, and constraints on 43 linear combinations of WCs, are obtained.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

SV-Sim: Scalable PGAS-based State Vector Simulation of Quantum Circuits

High-performance quantum circuit simulation in a classic HPC is still imperative in the NISQ era. Observing that the major obstacle of scalable state-vector quantum simulation arises from the massively fine-grained irregular data-exchange with remote nodes, in this paper we present SV-Sim to apply the emerging PGAS-based communication models (i.e., direct peer access for intra-node CPUs/GPUs and SHMEM for inter-node CPU/GPU clusters) for efficient scalable quantum circuit simulation. Through an orchestrated device functional pointer design, SV-Sim is able to abstract the quantum gate sets across various heterogeneous backends, including IBM/Intel/AMD CPUs, NVIDIA /AMD GPUs, and Intel MIC, in a unified framework, but still asserting outstanding performance and tractable interface to higher-level quantum programming environments, such as IBM Qiskit, Microsoft Q\# and Google Cirq. Circumventing the disability of polymorphism in GPUs and leveraging the device-initiated one-sided communication, SV-Sim can process dynamically synthesized quantum circuit in a single GPU/CPU kernel without the need of expensive JIT or runtime branching, significantly improving the performance and simplifying the programming complexity for the emerging variational quantum algorithms. Evaluations on NVIDIA A100-DGX-1, V100-DGX-2, AMD MI100, ALCF Theta, and OLCF Summit HPCs show that SV-Sim can delivery scalable performance on various state-of-the-art HPC platforms, offering a useful tool for quantum algorithm validation and verification.

Li, Ang↗

A Study of Performance Portability of Low-bit Fused Matrix-Vector Multiplication Kernels in SYCL

Understanding the causes of performance gaps between a portable programming model and a vendor-specific programming model is important for improving performance portability. This paper studies performance portability of low-bit fused general matrix-vector multiplication kernels in SYCL on vendors’ graphics processing units (GPUs). This work introduces the use case, explains the kernel implementations in detail, evaluates the performance of the CUDA, HIP, and SYCL kernels on datacenter, desktop, and laptop GPUs, and investigates the causes of performance gaps. The results show that loop unrolling, kernel dispatch overhead, and sum reduction contribute to the gaps.

Jin, Zheming [ORNL] (ORCID:000000027197780X)↗