Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Mesh Optimization”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

DISTRI: Distributed Multi-Facility HPC Simulator (DISTRI) v2.1

DISTRI is an advanced network simulator designed for multi-facility computational infrastructures with agentic behavior. It simulates HPC facilities where computational resources act as autonomous agents, making intelligent decisions about job scheduling, load balancing, and resource allocation. The simulator focuses on developing and testing decentralized algorithms that promote resilience and efficiency in multi-facility environments. Key Features: - Agentic Resource Behavior: Processors and DTNs act as autonomous agents with decision-making capabilities - Pheromone-Based Load Balancing: Decentralized load balancing inspired by ant colony optimization - Dual Topology Support: Mesh (normal operations) and Dumbell (network testing) topologies - Comprehensive TCP Simulation: Realistic TCP implementations with multiple congestion control algorithms - Failure Resilience Testing: Processor failure simulation with automatic job reassignment - Extensive Visualization: Detailed performance analysis and metrics collection - Research-Ready: Designed for algorithm development and benchmarking

Bez, Jean Luca [Lawrence Berkeley National Laborat↗

Analysis of cell-based diffusion acceleration for the slice balance approach

In this work, we perform analysis on the use of cell-based diffusion acceleration methodologies to accelerate the convergence of transport solutions discretized with the slice balance approach (SBA) on unstructured polygonal grids.We investigated both linear diffusion synthetic acceleration (DSA) and non linear diffusion acceleration (NDA), including its partial-current variant (pNDA). DSA and NDA were both shown to diverge for intermediate ranges of mesh optical thicknesses. However, pNDA and Krylov methods like GMRES and Broyden stabilized the acceleration schemes, including problems with degenerate cells formed by mesh refinement. (author)

42 ENGINEERING↗

Parallel transport sweeps on two-dimensional cartesian and hexagonal grids

This paper aims to provide a proof of concept for parallel transport sweeps on two-dimensional hexagonal grids for the discrete ordinates transport equation. While the method is an extension of the popular and well-established Koch-Baker-Alcoulffe (KBA) algorithm, there are significant differences between the cartesian and hexagonal grid and thereafter sweep. The most important is the three-way connectivity of hexagons within the grid which creates greater dependencies between the elements. The KBA method in structured orthogonal grids was first implemented in the DRAGON5 code and the method is first described here. The differences in implementation for the hexagonal grid are also described. Benchmark results are also presented, showing roughly 10 times speedup in computational times with roughly 100 processors, in both cases. (authors)

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Application of linear prolongation to coarse mesh finite difference acceleration in CASMO5

The Coarse-Mesh Finite Difference (CMFD) method has been used for over a decade to accelerate the convergence of the Method of Characteristics (MOC) solution to the two- dimensional particle transport equation in CASMO5. Numerical testing, along with widespread use in production-level calculations, have shown that the current CMFD implementation provides stability and robustness for a wide range of realistic reactor physics problems. However, the recent development of linear prolongation has attracted attention from the community as a way to further improve the performance and stability of CMFD. Two interpolation methods for linear prolongation are presented in this work and implemented into a test version of CASMO5. The performance of the proposed interpolations, referred to as the bilinear and linear directional schemes, is evaluated in terms of runtime relative to the default constant or uniform scaling update. Numerical results indicate that the use of linear prolongation can reduce the transport solver runtime on average by approximately 10% when tested with two hundred randomly selected cases. The new directional linear interpolation, combined with default constant boundary updates, is found to provide the highest reduction in runtime for the cases analyzed. (authors)

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Implementation of a Mesh refinement algorithm into the quasi-static PIC code QuickPIC

Plasma-based acceleration (PBA) has emerged as a promising candidate for the accelerator technology used to build a future linear collider and/or an advanced light source. In PBA, a trailing or witness particle beam is accelerated in the plasma wave wakefield (WF) created by a laser or particle beam driver. The WF is often nonlinear and involves the crossing of plasma particle trajectories in real space and thus particle-in-cell methods are used. The distance over which the drive beam evolves is several orders of magnitude larger than the wake wavelength. This large disparity in length scales is amenable to the quasi-static approach. Three-dimensional (3D), quasi-static (QS), particle-in-cell (PIC) codes, e.g., QuickPIC, have been shown to provide high fidelity simulation capability with 2-4 orders of magnitude speedup over 3D fully explicit PIC codes. In PBA, the witness beam needs to be matched to the focusing forces of the WF to reduce the emittance growth. In some linear collider designs, the matched spot size of the witness beam can be 2 to 3 orders of magnitude smaller than the spot size (and wavelength) of the wakefield. Such an additional disparity in length scales is ideal for mesh refinement where the WF within the witness beam is described on a finer mesh than the rest of the WF. A mesh refinement scheme is described that has been implemented into the 3D QS PIC code, QuickPIC. Very fine (high) resolution is used in a small spatial region that includes the witness beam and progressively coarser resolutions in the rest of the simulation domain. A fast multigrid Poisson solver has been implemented for the field solve on the refined meshes and a Fast Fourier Transform (FFT) based Poisson solver is used for the coarse mesh. The code has been parallelized with both MPI and OpenMP, and the parallel scalability has also been improved by using pipelining. A preliminary adaptive mesh refinement technique is described to optimize the computational time for simulations with an evolving witness beam size. Several test problems are used to verify that the mesh refinement algorithm provides accurate results. Additionally, the results are benchmarked against highly resolved simulations exhibiting near-azimuthal symmetry, performed using QPAD—a novel hybrid QS PIC code that uses a PIC description in the coordinates (r, ct – z) and a gridless description in the azimuthal angle, Φ.

Linear collider↗

Optimization of the moderators in the STS preliminary design

This report details the results for an optimization of the dimensions of the moderators in the preliminary design of the Spallation Neutron Source Second Target Station (STS). This study uses the optimization algorithms of Dakota and an unstructured mesh model for the moderators in MCNP. More details on the unstructured mesh model and the automated mesh generation can be found in [3]. Parallel to this effort, the same moderator geometries have been optimized using a constructive solid geometry (CSG) MCNP model. More details on this model and its results can be found in [4]. Three optimal designs are selected for each moderator: one that is optimized for maximum peak brightness, one for maximum time-integrated brightness, and one for a combination of peak and time-integrated brightness. The backbone of the optimization work flow is provided by Dakota. For each set of design parameters requested by Dakota, a new solid geometry is automatically built in Creo and SpaceClaim, and subsequently exported to Attila4MC to generate an unstructured mesh geometry for MCNP. After the MCNP calculation is finished, the objective function (e.g., brightness metric) is returned to Dakota. After the new design has been evaluated, a result-file is written, and Dakota proposes the next set of design parameters to be evaluated. The loop continues until a specified convergence criterion has been met. The design parameters of the cylindrical (upper) moderator include the hydrogen radius, the premoderator thickness (top, bottom, radial), the beryllium radius and the horizontal position of the moderator. The crucial design choice is the hydrogen radius. A radius of 62 mm is shown to provide the maximum time-integrated brightness. The maximum peak brightness occurs with a radius of 40 mm. A combined (middle) design, which balances peak and time-integrated brightnesses, is obtained with a hydrogen radius of 50 mm. The premoderator thicknesses and the beryllium radius are slightly larger in the design optimized for time-integrated brightness than in the design optimized for peak brightness. The sensitivity to these two parameters is relatively small close to the optimal configurations. The hydrogen vessel and vacuum vessel wall thicknesses are dependent on the radius of the liquid hydrogen due to structural integrity requirements. The increased wall thicknesses for larger vessels significantly penalize the time-integrated brightness, with the maximum obtainable value reduced by more than 10% relative to earlier studies which used fixed vessel wall thicknesses. The impact of the variable wall thicknesses is much less for the peak brightness and combined brightness designs. The design parameters of the tube (lower) moderator selected for the optimization are the tube length, the annular premoderator thickness, the beryllium radius and the horizontal position of the moderator. The tube length is the crucial parameter and is chosen large (210 mm) and small (125 mm) in the designs optimized for time-integrated and peak brightness respectively. A combined optimal design has a tube length of 170 mm. The premoderator thickness and the beryllium radius are chosen larger in the design optimized for time-integrated brightness.

42 ENGINEERING↗

Examination of a Methane/Diesel RCCI Engine Using Pele: Preprint

Multi-fuel, advanced injection strategies have become increasingly promising as a strategy to mitigate the emissions generated from internal combustion engines. By carefully controlling the combustion phasing in-cylinder, these new multi-pulse, multi-fuel injection strategies are able to burn in the low-temperature combustion regime where both NOx and soot are not readily produced, reducing the need for extensive exhaust gas recirculation systems. In this study, we examine a reactivity-controlled compression ignition (RCCI) strategy that uses an early pre-filled methane-air mixture with low turbulence background as the low-reactivity fuel and a direct injection of four discrete dodecane jets as a surrogate for the high-reactivity diesel fuel. We use the Pele software suite, a highly optimized, exascale-ready, adaptive mesh refinement codebase to perform high-resolution numerical simulations of a scaled down, single cylinder from the RCCI engine. Here, we resolve the ignition kernels down to micrometer scales and present several statistical quantities evaluating the development of the flow and detailing the onset of ignition and subsequent flame development. Particular attention is paid to the conditions surrounding the onset of the first ignition kernels and discussing what led to the development of those conditions.

CFD↗

Enabling the broader adoption of fusion simulation on complex geometry

This project addressed a key barrier to advanced fusion and nuclear simulation: the difficulty of performing high-fidelity Monte Carlo neutronics directly on complex, real-world CAD geometry. Traditional workflows require engineers to rebuild CAD models as simplified constructive solid geometry, a time-consuming and error-prone process that limits design iteration and broader adoption of simulation tools. The goal of this Phase I SBIR was to make CAD-based neutronics practical, accessible, and robust for industrial and research users. During the project, Coreform significantly enhanced the Direct Accelerated Geometry Monte Carlo (DAGMC) workflow and fully integrated it into Coreform Cubit as a first-class capability. Major achievements include optimized material assignment and surface meshing workflows, substantial performance improvements to geometry imprinting and preparation, native export of DAGMC models, and new visualization tools to support OpenMC source definition and lost-particle debugging. Coreform also expanded Cubit’s capabilities as a full OpenMC preprocessor, including the ability to convert OpenMC constructive solid geometry models back into CAD for visualization, multiphysics coupling, and debugging. In collaboration with Argonne National Laboratory, the project delivered comprehensive new DAGMC documentation and training materials, transforming DAGMC from a research-oriented tool into a production-ready workflow. Results were disseminated through tutorials, conference training, and multiple well-attended webinars demonstrating integrated CAD-based neutronics and multiphysics workflows. Overall, this project demonstrated that high-fidelity Monte Carlo simulations can be performed directly on complex CAD geometry, reducing setup time, improving usability, and enabling faster, more informed design decisions for fusion and nuclear energy systems.

42 ENGINEERING↗

Optimal local truncation error method for 3-D elasticity interface problems

The paper deals with a new effective numerical technique on unfitted Cartesian meshes for simulations of heterogeneous elastic materials. Here, we develop the optimal local truncation error method (OLTEM) with 27- point stencils (similar to those for linear finite elements) for the 3-D time-independent elasticity equations with irregular interfaces. Only displacement unknowns at each internal Cartesian grid point are used. The interface conditions are added to the expression for the local truncation error and do not change the width of the stencils. The unknown stencil coefficients are calculated by the minimization of the local truncation error of the stencil equations and yield the optimal second order of accuracy for OLTEM with the 27-point stencils on unfitted Cartesian meshes. A new post-processing procedure for accurate stress calculations has been developed. Similar to basic computations it uses OLTEM with the 27-point stencils and the elasticity equations. The post-processing procedure can be easily extended to unstructured meshes and can be independently used with existing numerical techniques (e.g., with finite elements). Numerical experiments show that at an accuracy of 0.1% for stresses, OLTEM with the new post-processing procedure significantly (by 10 5 -10 9 times) reduces the number of degrees of freedom compared to linear finite elements. OLTEM with the 27-point stencils yields even more accurate results than high-order finite elements with wider stencils.

42 ENGINEERING↗

R-Adaptivity to Enable Compression of Elementary Computations in Extreme-Scale Finite Element Simulators

Modern computing systems are capable of exascale calculations, which are revolutionizing the development and application of high-fidelity numerical models in computational science and engineering. While these systems continue to grow in processing power, the available system memory has not increased commensurately, and electrical power consumption continues to grow. A predominant approach to limit the memory usage in large-scale applications is to exploit the abundant processing power and continually recompute many low-level simulation quantities, rather than storing them. However, this approach can adversely impact the throughput of the simulation and diminish the benefits of modern computing architectures. We present three novel contributions to reduce the memory burden while maintaining, and sometimes improving, performance in simulations based on finite element discretizations. The first contribution develops dictionary-based data compression schemes that detect and exploit the structure of the discretization, due to redundancies across the finite element mesh. While these schemes are shown to reduce memory requirements by more than 99% on meshes with large numbers of identical mesh cells, there are applications where this structure does not exist. The second contribution leverages a recently developed augmented Lagrangian optimization algorithm to enable r-adaptivity for meshes with the goal of enhancing the redundancies in the mesh. The third contribution extends these methods to patch-based linear solvers and preconditioners by compressing local matrices. Numerical results demonstrate the effectiveness of the proposed methods to detect, enhance and exploit mesh structure on a suite of examples inspired by large-scale applications.

97 MATHEMATICS AND COMPUTING↗

A method for bounding high-order finite element functions: Applications to mesh validity and bounds-preserving limiters

We introduce a novel method for bounding high-order multi-dimensional polynomials in finite element approximations. The method involves precomputing optimal piecewise-linear bounding boxes for polynomial basis functions, which can then be used to locally bound any combination of these basis functions. This approach can be applied to any element/basis type at any approximation order, can provide local (i.e., subcell) extremum bounds to a desired level of accuracy, and can be evaluated efficiently on-the-fly in simulations. Furthermore, we show that this approach generally yields more accurate bounds in comparison to traditional methods based on convex hull properties (e.g., Bernstein polynomials). Furthermore, the efficacy of this technique is shown in applications such as mesh validity checks and optimization for high-order curved meshes, where positivity of the element Jacobian determinant can be ensured throughout the entire element, and continuously bounds-preserving limiters for hyperbolic systems, which can enforce maximum principle bounds across the entire solution polynomial.

Bounding box↗

Development of a self-lubricating high-efficiency hybrid seal composed of carbon nanotube-coated metal meshes for CSP turbomachinery (SETO CPS #36333 Final Report)

In turbomachinery, internal leakage flow accounts for up to 3% of the total thermodynamic cycle energy loss. Tradeoff must be made between the sealing efficiency (smaller clearance) and the friction and wear issues for interfering with the shaft (larger clearance). This ORNL-Danfoss joint effort developed a novel hybrid seal composed of carbon nanotube (CNT)-coated metal meshes. The CNT growth process was based on a self-catalyzing chemical vapor deposition and these multiwall CNTs were well aligned with high crystallinity. This hybrid material structure takes advantage of the CNT’s low-friction nature and uses the metal mesh as an extendable backbone. Full-scale experimental seals were designed, fabricated, and optimized for sealing performance and durability. The CNT-coated metal mesh seal demonstrated superior gas sealing efficiency to the baseline labyrinth seal and significantly improved shaft surface protection compared with the state-of-the-art superalloy brush seal on the static rig and full-scale compressor dynamometer tests. The CNT-metal mesh seal is low-cost and scalable and can potentially benefit wide applications, including CSP and other power generation, marine, automotive, and HVAC.

36 MATERIALS SCIENCE↗

Coupled Target-Beam-Moderator Optimization for the Second Target Station

This report describes the results for a coupled target-beam-moderator optimization analysis for the Second Target Station (STS) at ORNL's Spallation Neutron Source. This study is a continuation of the optimization analysis for the moderators in the preliminary design of STS performed in 2022. In the 2022 analysis the dimensions of the moderators are parameterized, while the target and the proton beam profile are kept constant. In this analysis the target height and the proton beam profile are added as parameters. This allows to study the coupled effects of changing target, moderator and beam dimensions. Similar to the 2022 analysis, this work is performed with an automated optimization workflow that uses the optimization toolbox DAKOTA, parameterized geometries in CREO and SpaceClaim, the unstructured mesh generation in Attila4MC, and the particle transport code MCNP6.2©. This workflow enables an efficient optimization using high-fidelity geometries. The main conclusions of this analysis are the following: • Coupled beam-target-moderator optimization provides a few additional percent performance gain over stand-alone moderator optimization. • The moderator performance is not very sensitive to the target height (between ≈60 and ≈80 mm) as long as the beam profile is chosen adequately. • The moderator performance is sensitive to the choice of beam spatial standard deviations, even when the footprint is kept constant. • The optimal moderator radius is the same for a beam footprint of 30 cm 2 , 62.5 cm 2 , and 90 cm 2 . Also the slope of the super-gaussian beam profile does not significantly impact the optimal moderator radius. • The optimal parameters and sensitivities are very similar to the 2022 optimization analysis. These results only indicate a a difference in the optimal radius of the cylindrical moderator, however, this has been corrected in the final design moderator optimization. The main purpose of this report is to document the simulations, results and lessons learned. The most impactful results are summarized in. We also note that the target geometry used in this work is not the final design.

43 PARTICLE ACCELERATORS↗

Energy Optimization of Light and Heavy-Duty Vehicle Cohorts of Mixed Connectivity, Automation and Propulsion System Capabilities via Meshed V2V-V2I and Expanded Data Sharing (Final Scientific and Technical Report)

Vehicle connectivity and automated driving technologies individually have the potential to decrease energy consumption and/or increase safety on light, medium or heavy duty vehicles to varying degrees depending on the traffic infrastructure and specific driving scenarios. Due to advances in sensing, perception and computing power, research and development emphasis in the mobility sector has shifted away from connectivity. Prior research has shown that driving automation with the absence of connectivity can in certain circumstances increase energy consumption. The effectiveness of synergizing connectivity and driving automation technologies is the focus of this work, specifically applied to vehicle cohorts of mixed composition, light and heavy duty, and powertrains ranging from all electric to conventional internal combustion engine. The project team is led by Michigan Technological University (MTU) and partnered with AVL Mobility Technologies Inc. (AVL), Borg Warner (BW), Traffic Technology Services (TTS), American Center for Mobility (ACM) and Navistar (NAV). The main thrusts for the team are to develop a micro-traffic simulation environment with specific VD&PT system attributes and CAV capabilities, 2) field a vehicle test fleet of mixed classification, propulsion and CAV capacity, 3) develop artificial intelligence (AI) and machine learning (ML) based multi-agent optimization methods for various traffic infrastructures, 4) integrate the virtual environment and the optimization methods then deploy the system as a CAV hardware in the loop (HiL) for the vehicle test fleet and 5) conduct closed track and public road testing to validate simulation and demonstrated energy and mobility improvements at multiple scales. For a cohort of mixed vehicles, the team will demonstrate a reduction of energy consumption of 10-50% at intersection, arterial roadway and limited access highway scenarios through connectivity and automation in simulation and at a closed test track. The energy reduction objectives of the project are summarized in Table 1, indicating the infrastructure and over what distances are relevant considered. Single scenario energy reductions are not relevant and thus, the research team took the approach to vary parameters associated with the infrastructure, vehicle cohort composition and dynamic behavior to generate energy consumption distributions for both unconnected and connected scenarios.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Gaussian integral method for void fraction

Here, a novel method, the Gaussian Integral Method (GIM), is presented for calculating void fractions in Computational Fluid Dynamics–Discrete Element Method (CFD-DEM) simulations. GIM is versatile and applicable to various grid types, including structured and unstructured polyhedral meshes, without requiring special boundary treatments. An optimization technique is introduced to make GIM independent of grid resolution and type. The method is validated against experimental data from a fluidized bed, demonstrating that GIM produces realistic simulations closely resembling experimental observations. Additionally, unstructured polyhedral grids using GIM outperform structured grids of equivalent resolution, yielding results more aligned with experimental data. The gradient of the void fraction is computed in the CFD solver and utilized in the DEM solver for precise estimation at particle locations. Overall, GIM provides an effective solution for void fraction calculations in particulate media simulations with complex geometries, enhancing the accuracy and applicability of CFD-DEM simulations for industrial processes.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Design of proton deflectometry with in situ x-ray fiducial for magnetized high-energy-density systems

We report a design and implementation of proton deflectometry with an in situ reference x-ray image of a mesh to precisely measure non-uniform magnetic fields in expanding plasmas at the OMEGA and OMEGA EP laser facilities. The technique has been developed with proton and x-ray sources generated from both directly driven capsule implosions and short pulse laser–solid interactions. The accuracy of the measurement depends on the contrast of both the proton and x-ray images. Here we present numerical and analytic studies to optimize the image contrast using a variety of mesh materials and grid spacings. Our results show clear enhancement of the image contrast by factors of four to six using a high Z mesh with large grid spacing. This leads to further improvement in the accuracy of the magnetic field measurement using this technique in comparison with its first demonstration at the OMEGA laser facility [Rev. Sci. Instrum. 93, 023502 (2022) [CrossRef]].

47 OTHER INSTRUMENTATION↗