Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “parallelization”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 361 records · Page 20

LibERI—A portable and performant multi-GPU accelerated library for electron repulsion integrals via OpenMP offloading and standard language parallelism

A portable and performant graphics processing unit (GPU)-accelerated library for electron repulsion integral (ERI) evaluation, named LibERI, has been developed and implemented via directive-based (e.g., OpenMP and OpenACC) and standard language parallelism (e.g., Fortran DO CONCURRENT). Offloaded ERIs consist of integrals over low and high contraction s, p, and d functions using the rotated-axis and Rys quadrature methods. GPU codes are factorized based on previous developments with two layers of integral screening and quartet presorting. In this work, the density screening is moved to the GPU to enhance the computational efficacy for large molecular systems. Here, the L-shells in the Pople basis set are also separated into pure S and P shells to increase the ERI homogeneity and reduce atomic operations and the memory footprint. LibERI is compatible with any quantum chemistry drivers supporting the MolSSI Driver Interface. Benchmark calculations of LibERI interfaced with the GAMESS software package were carried out on various GPU architectures and molecular systems. The results show that the LibERI performance is comparable to other state-of-the-art GPU-accelerated codes (e.g., TeraChem and GMSHPC) and, in some cases, outperforms conventionally developed ERI CUDA kernels (e.g., QUICK) while fully maintaining portability.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Nonlinear Poisson–Boltzmann solutions for charged parallel plates: When opposite charges repel

I present an exact solution of the Poisson–Boltzmann equation for two parallel plates and discuss the solution properties. I discuss in more detail plates with opposite charges: In this case, there are two critical separations, L c,1 < L c,2 . For separations less than L c,1 , the force between plates is repulsive. It switches to attractive at L c,1 , but with the electric potential having the same sign on both plates. For L > L c,2 , the force remains attractive, and the potential at the plates has the same sign as the charge on each plate. I also describe charge regulation, determined by pK a , and provide formulas for both the critical distance where oppositely charged plates repel and their charging process. Finally, the implications of these results for the nanoparticle assembly, as driven by electrostatic interactions, are also discussed.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Mechanical design of a parallel flexure-based RADSI instrument for curved x-ray mirror metrology

Modern synchrotron x-ray beamlines demand reflective optics with higher surface profile accuracy to achieve diffraction-limited focusing. This necessitates advanced metrology instruments capable of delivering repeatable measurements in the nanometer to sub-nanometer range. Slope ranges exceeding 15 mrad (0.86°) and greater pose significant challenges for mirror metrology using conventional interferometric methods. Here, to address this, we present a new relative angle determinable stitching interferometry instrument featuring a parallel flexure-based mechanical design. This approach enhances vibration and thermal stability while maintaining a compact and lightweight system. Initial measurements of a cylindrical mirror with a 16 m radius of curvature and a slope range of 5 mrad demonstrate nanometer-level repeatability. Comprehensive system characterization suggests the potential for achieving sub-nanometer repeatability with further refinement to the instrument.

36 MATERIALS SCIENCE↗

Saturation of the kinetic ballooning instability due to the electron parallel nonlinearity

The electron parallel nonlinearity (EPN) is implemented in the gyrokinetic particle-in-cell turbulence code GEM [Y. Chen and S. E. Parker, J. Comp. Phys. 220, 839 (2007)]. Application to the Cyclone Base Case reveals a strong effect of EPN on the saturated heat transport above the kinetic ballooning mode (KBM) threshold. Evidence is provided to show that the strong effect is associated with the electron radial motion due to magnetic fluttering, which turns fine structures of the KBM eigenmode in radius into fine structures in velocity and increases the magnitude of the EPN term in the kinetic equation.

Gyrokinetic simulations↗

Massively parallel, computationally guided design of a proenzyme

Proteins have shown promise as therapeutics and diagnostics, but their effectiveness is limited by our inability to spatially target their activity. To overcome this limitation, we developed a computationally guided method to design inactive proenzymes or zymogens, which are activated through cleavage by a protease. Since proteases are differentially expressed in various tissues and disease states, including cancer, these proenzymes could be targeted to the desired microenvironment. We tested our method on the therapeutically relevant protein carboxypeptidase G2 (CPG2). We designed Pro-CPG2s that are inhibited by 80 to 98% and are partially to fully reactivatable following protease treatment. The developed methodology, with further refinements, could pave the way for routinely designing protease-activated protein-based therapeutics and diagnostics that act in a spatially controlled manner. Confining the activity of a designed protein to a specific microenvironment would have broad-ranging applications, such as enabling cell type-specific therapeutic action by enzymes while avoiding off-target effects. While many natural enzymes are synthesized as inactive zymogens that can be activated by proteolysis, it has been challenging to redesign any chosen enzyme to be similarly stimulus responsive. Here, we develop a massively parallel computational design, screening, and next-generation sequencing-based approach for proenzyme design. For a model system, we employ carboxypeptidase G2 (CPG2), a clinically approved enzyme that has applications in both the treatment of cancer and controlling drug toxicity. Detailed kinetic characterization of the most effectively designed variants shows that they are inhibited by ∼80% compared to the unmodified protein, and their activity is fully restored following incubation with site-specific proteases. Introducing disulfide bonds between the pro- and catalytic domains based on the design models increases the degree of inhibition to 98% but decreases the degree of restoration of activity by proteolysis. A selected disulfide-containing proenzyme exhibits significantly lower activity relative to the fully activated enzyme when evaluated in cell culture. Structural and thermodynamic characterization provides detailed insights into the prodomain binding and inhibition mechanisms. The described methodology is general and could enable the design of a variety of proproteins with precise spatial regulation.

59 BASIC BIOLOGICAL SCIENCES↗

Influences of δB contribution and parallel inertial term of energetic particles on MHD-kinetic hybrid simulations: a case study of the 1/1 internal kink mode

The magnetohydrodynamic-kinetic (MHD-kinetic) hybrid model (Park et al 1992 Phys. Fluids B 4 2033–7) has been widely applied in studying energetic particles (EPs) problems in fusion plasmas for past decades. The pressure-coupling scheme or the current-coupling scheme is adopted in this model. However, two noteworthy issues arise in the model application: firstly, the coupled term introduced in the pressure-coupling scheme, (∇•P h ) ⟂ , is often simplified by ∇•P h , which is equivalent to neglecting the parallel inertial term of EPs; secondly, besides the $δf$ contribution caused by changing in the EP distribution function, the magnetic field perturbation (the $δB$ contribution) generated during development of the instabilities should also be considered, but it is often ignored in existing hybrid simulations. In this paper, we derive the analytical formulations under these two coupling schemes and then numerically study the representative case of the linear stability of the $m/n$ = $1/1$ internal kink mode (IKM) (Fu et al 2006 Phys. Plasmas 13 052517) by using the CLT-K code. Further, it is found that the approximated models can still yield reasonable results when EPs are isotopically distributed. But it fails completely in cases with anisotropic EP distributions. In addition, we further investigate the influence of EP's orbit width on the stability of IKM and verify the equivalence between pressure-coupling scheme and the current-coupling scheme.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

2D measurements of parallel counter-streaming flows in the W7-X scrape-off layer for attached and detached plasmas

Abstract Investigations of particle parallel flow velocities have been carried out for the scrape-off layer (SOL) of the Wendelstein 7-X (W7-X) stellarator, in order to gain insights on the SOL transport properties during attached and detached plasma scenarios. The experimental evidence is based on the coherence imaging spectroscopy (CIS) diagnostic, able to measure 2D impurity emission intensity and flow velocity. The impurity monitored by CIS is C 2+ , characterized by a line-emission intensity observed to be linearly proportional to the total plasma radiated power in both attached and detached plasmas. The related C 2+ velocity shows a strong dependence on the line-averaged electron density while remaining insensitive to the input power. During attached plasmas, the velocity increases with increasing line-averaged density. The tendency reverses in the transition to and during detachment, in which the velocity decreases by at least a factor of 2. The sharp drop in velocity, together with a rise in line-emission intensity, is reliably correlated to the detachment transition and can therefore be used as one of its signatures. The impurity flow velocity appears to be well coupled with the main ions’ one, thus implying the dominant role of impurity-main ion friction in the parallelimpurity transport dynamics. In view of this SOL impurity transport regime, the CIS measurement results are here interpreted with the help of EMC3-Eirene simulations, and their major trends are already explainable with a simple 1D fluid model.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Parallel transport modeling of linear divertor simulators with fundamental ion cyclotron heating *

Abstract The Material Plasma Exposure eXperiment (MPEX) is a steady state linear device with the goal to perform plasma material interaction studies at future fusion reactor relevant conditions. A prototype of MPEX referred as ‘Proto-MPEX’ is designed to carry out research and development related to source, heating and transport concepts on the planned full MPEX device. The auxiliary heating schemes in MPEX are based on cyclotron resonance heating with radio frequency (RF) waves. Ion cyclotron heating (ICH) and electron cyclotron heating in MPEX are used to independently heat the ions and electrons and provide fusion divertor conditions ranging from sheath-limited to fully detached divertor regimes at a material target. A hybrid particle-in-cell code- PICOS++ is developed and applied to understand the plasma parallel transport during ICH in MPEX/Proto-MPEX to the target. With this tool, evolution of the distribution function of MPEX/Proto-MPEX ions is modeled in the presence of (a) Coulomb collisions, (b) volumetric particle sources and (c) quasi-linear RF-based ICH. The code is benchmarked against experimental data from Proto-MPEX and simulation data from B2.5 EIRENE. The experimental observation of ‘density-drop’ near the target in Proto-MPEX and MPEX during ICH is demonstrated and explained via physics-based arguments using PICOS++ modeling. In fact, the density drops at the target during ICH in Proto-MPEX/MPEX to conserve the flux and to compensate for the increased flow during ICH. Furthermore, sensitivity scans of various plasma parameters with respect to ICH power are performed for MPEX to investigate its role on plasma transport and particle and energy fluxes at the target.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Integrable multistate Landau–Zener models with parallel energy levels

In this paper, we discuss solvable multistate Landau–Zener (MLZ) models whose Hamiltonians have commuting partner operators with ~1/τ-time-dependent parameters. Many already known solvable MLZ models belong precisely to this class. We derive the integrability conditions on the parameters of such commuting operators, and demonstrate how to use such conditions in order to derive new solvable cases. We show that MLZ models from this class must contain bands of parallel diabatic energy levels. The structure of the scattering matrix and other properties are found to be the same as in the previously discussed completely solvable MLZ Hamiltonians.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Casimir interactions between two parallel graphene sheets carrying steady-state drift currents

Here, we investigate the fluctuation-induced Casimir interactions between two parallel graphene sheets carrying steady-state drift currents. The graphene properties are modeled based on the shifted Fermi disk model to capture the nonequilibrium optical response of the system. We find that the drift current introduces a repulsive correction to the perpendicular to the layers Casimir interaction, thereby reducing the overall attractive force. Although the correction is repulsive, it does not overcome the underlying attraction between the layers. It also generates a lateral force that opposes the carrier flow direction. Both contributions are studied in terms of distance and drift velocity functionalities showing pathways for Casimir force control.

Ke, Modi [University of South Florida, Tampa, FL (↗

Parallel-plate avalanche counters for heavy-ion beam tracking: History and mysteries

Despite being an old detector concept, position-sensitive parallel-plate avalanche counters (PPACs) remain widely used today for heavy-ion position and timing measurements. In modern rare isotope beam facilities and large-acceptance fragment separators, PPACs are used to characterize beam properties (diagnostics), facilitate beam delivery (tuning), and provide event-by-event beam tracking for particle identification (PID). Most popular particle localization methods in PPAC detectors are based on strip electrodes electrically connected to resistive-chain circuits or delay lines. More exotic systems include optical readouts based on recording electroluminescence light with high-granularity photodetector arrays or high-resolution resistive anode electrodes. PPACs with conventional resistive-chain or delay-line readouts achieve typical position and time resolutions of below 1 mm ( σ ) and around 150 ps ( σ ), respectively. In addition, delay-line systems are capable of counting rates above several hundred kHz for beam areas of a few millimeters square, and around 1 MHz for larger beam sizes. Resistive-chain readouts have limited rates of a few tens of kHz. A review of the operation principles and performance of PPAC detectors is presented in this paper. We will discuss decades-long experience building and operating PPACs developed at the National Superconducting Cyclotron Laboratory, and then refined at the Facility for Rare Isotope Beams (FRIB), mostly focusing on the delay-line readout technique (DLPPAC). We will also discuss problems that arise due to electric discharges at high rates and briefly describe ongoing developments at FRIB. Published by the American Physical Society 2024

Physics↗

Isotropic parallel antiferromagnetism in the magnetic field induced charge-ordered state of $\mathrm{Sm Ru_4 P_{12}}$ caused by $p - f$ hybridization

Nature of the field-induced charge-ordered phase (phase II) of SmRu 4 P 12 has been investigated by resonant x-ray diffraction (RXD) and polarized neutron diffraction (PND), focusing on the relationship between the atomic displacements and the antiferromagnetic (AFM) moments of Sm. From the analysis of the interference between the nonresonant Thomson scattering and the resonant magnetic scattering, combined with the spectral function obtained from x-ray magnetic circular dichroism, it is shown that the AFM moment of Sm prefers to be parallel to the field (m AF ∥ H), giving rise to large and small moment sites around which the P 12 and Ru cage contract and expand, respectively. This is associated with the formation of the staggered ordering of the Γ 7 -like and Γ 8 -like crystal-field states, providing a strong piece of evidence for the charge order. PND was also performed to obtain complementary and unambiguous conclusion. In addition, isotropic and continuous nature of phase II is demonstrated by the field-direction invariance of the interference spectrum in RXD. Finally, crucial role of the p-f hybridization is shown by resonant soft x-ray diffraction at the P K edge (1s↔3p), where we detected a resonance due to the spin polarized 3p orbitals reflecting the AFM order of Sm.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Effects of scattering on the field-induced T c enhancement in thin superconducting films in a parallel magnetic field

The problem of the normal-superconducting phase boundary for films in a parallel magnetic field, discussed in the classical paper by Ginzburg and Landau for temperatures close to the critical, is revisited with the help of the microscopic BCS theory for arbitrary temperatures taking pair-breaking and transport scattering into account. Although confirming experimental findings of the T c enhancement by the magnetic field, we find that the transport scattering pushes the phase transition curve to higher fields and higher temperatures for nearly all practical scattering rates. Still, the T c enhancement disappears in the dirty limit. Here, we also consider intriguing changes, such as reentrant superconductivity, caused to the phase boundary by pair-breaking magnetic ions spread on one of the film faces. These features await experimental verification.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Nature of quantum spin liquids of the S = 1 2 Heisenberg antiferromagnet on the triangular lattice: A parallel DMRG study

Here we study the ground-state properties of the quantum spin liquid (QSL) phases of the spin-1/2 antiferromagnetic Heisenberg model on the triangular lattice with nearest- (J 1 ), next-nearest- (J 2 ), and third-neighbor (J 3 ) interactions by using density-matrix renormalization group (DMRG) method. By combining parallel DMRG with SU(2) spin rotational symmetry, we are able to obtain accurate results on large cylinders with length up to L x =48 and circumference L y =6–12. Our results suggest that the QSL phase of the J 1 -J 2 Heisenberg model is gapped which is characterized by the absence of gapless mode, short-range spin-spin and dimer-dimer correlations. In the presence of J 3 interaction, we find that a new critical QSL with a single gapless mode emerges. While both spin-spin and scalar chiral-chiral correlations are short-ranged, dimer-dimer correlations are quasi-long-ranged which decays as a power-law at long distances.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Automated and highly parallelized Bayesian optimization scheme for direct drive fusion experiments on OMEGA

Finding the optimal implosion design on existing experimental facilities for inertial confinement fusion requires an exhaustive search of the vast design parameter space. This is infeasible both with experiments and with simulations. Consequently, a large fraction of the experimentally realizable design space remains unexplored, and new design schemes are challenging to optimize in a reasonable time frame. On the OMEGA laser facility, predictive machine learning models have been developed to accurately forecast the result of an experiment using only inexpensive simulations and the large dataset of prior experimental data. However, the full design space remains vast enough to be unassailable with simple optimization techniques. Here we develop an automated and optimally parallel Bayesian optimization algorithm that can entirely optimize the target and pulse shape of a direct-drive ICF implosion under a given design paradigm. We use this algorithm to find a markedly improved design for the performance implosions on OMEGA that is predicted to hydroequivalently scale to ignition at 2.15 MJ.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Parallel-in-time quantum simulation via Page and Wootters quantum time

In the past few decades, researchers have created a veritable zoo of quantum algorithms by drawing inspiration from classical computing, information theory, and even from physical phenomena. Here, we present quantum algorithms for parallel-in-time simulations that are inspired by the Page and Wootters formalism. In this framework, and thus in our algorithms, the classical time variable of quantum mechanics is promoted to the quantum realm by introducing a Hilbert space of “clock” qubits that are then entangled with the “system” qubits. We show that our algorithms can compute temporal properties over 𝑁 different times of many-body systems by only using log⁡(𝑁) clock qubits. As such, we achieve an exponential trade-off between time and spatial complexities. In addition, we rigorously prove that the entanglement created between the system qubits and the clock qubits has operational meaning, as it encodes valuable information about the system’s dynamics. We also provide a circuit depth estimation of all the protocols, showing a running time advantage in computation times over traditional sequential-in-time algorithms. In particular, for the case when the dynamics are determined by the Aubry-Andre model, we present a hybrid method for which our algorithms have a depth that only scales as 𝒪⁡(log⁡(𝑁)⁢𝑛). As a by-product, we can relate the previous schemes to the problem of equilibration of an isolated quantum system, thus indicating that our framework enables a new dimension for studying dynamical properties of many-body systems.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

A high-temperature sample changer for parallelized in situ X-ray studies to efficiently explore reaction space

A sample environment for high-throughput X-ray scattering studies in transmission geometry to probe the mechanism and kinetics of moderate-temperature reactions in solution, molten fluxes and solids is described. This high-temperature sample changer enables efficient studies of reactions that are slow relative to the timescale of the X-ray scattering measurements by allowing up to 18 samples to be probed at the same temperature in parallel. This significantly enhances the throughput ofin situX-ray scattering studies as the sample changer effectively facilitates systematic studies that compare different reaction parameters (e.g.concentration, precursor, composition, additives), reference samples (e.g.background, pure precursors) and replicates (to demonstrate reproducibility) with enhanced consistency afforded by the quasi-simultaneous nature of the measurements. The large sample volumes, compared with those typically used for X-ray scattering measurements, are on a similar scale to those in the laboratory, making the results more directly comparable.

Chemistry↗

BeeSwarm: Enabling Parallel Scaling Performance Measurement in Continuous Integration for HPC Applications

Testing is one of the most important steps in software development–it ensures the quality of software. Continuous Integration (CI) is a widely used testing standard that can report software quality to the developer in a timely manner during development progress. Performance, especially scalability, is another key factor for High Performance Computing (HPC) applications. There are many existing profiling and performance tools for HPC applications, but none of these are integrated into CI tools. In this work, we propose BeeSwarm, an HPC container based parallel scaling performance system that can be easily applied to the current CI test environments. BeeSwarm is mainly designed for HPC application developers who need to monitor how their applications can scale on different compute resources. We demonstrate BeeSwarm using a multi-physics HPC application with Travis CI, GitLab CI and GitHub Actions while using ChameleonCloud and Google Compute Engine as the compute backends. Finally, our results show that BeeSwarm can be used for scalability and performance testing of HPC applications.

97 MATHEMATICS AND COMPUTING↗