Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “layout”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 127 records · Page 7

Accelerating Random Forest Classification on GPU and FPGA

Random Forests (RFs) are a commonly used machine learning method for classification and regression tasks spanning a variety of application domains, including bioinformatics, business analytics, and software optimization. While prior work has focused primarily on improving performance of the training of RFs, many applications, such as malware identification, cancer prediction, and banking fraud detection, require fast RF classification. In this work, we accelerate RF classification on GPU and FPGA. In order to provide efficient support for large datasets, we propose a hierarchical memory layout suitable to the GPU/FPGA memory hierarchy. We design three RF classification code variants based on that layout, and we investigate GPU- and FPGA-specific considerations for these kernels. Our experimental evaluation, performed on an Nvidia Xp GPU and on a Xilinx Alveo U250 FPGA accelerator card using publicly available datasets on the scale of millions of samples and tens of features, covers various aspects. First, we evaluate the performance benefits of our hierarchical data structure over the standard compressed sparse row (CSR) format. Second, we compare our GPU implementation with cuML, a machine learning library targeting Nvidia GPUs. Third, we explore the performance/accuracy tradeoff resulting from the use of different tree depths in the RF. Finally, we perform a comparative performance analysis of our GPU and FPGA implementations. Our evaluation shows that for high accuracy targets, our GPU implementation yields 5-9x speedup over CSR, and up to a 2x speedup over cuML.

FPGA, Xilinx FPGA, GPU, Random Forest classificati↗

Realistic Cost to Execute Practical Quantum Circuits using Direct Clifford+T Lattice Surgery Compilation

We report a resource estimation pipeline that explicitly compiles quantum circuits expressed using the Clifford+T gate set into a surface code lattice surgery instruction set. The cadence of magic state requests from the compiled circuit enables the optimization of magic state distillation and storage requirements in a post-hoc analysis. To compile logical circuits into lattice surgery operations, we build upon the open-source Lattice Surgery Compiler. The revised compiler operates in two stages: the first translates logical gates into an abstract, layout-independent instruction set; the second compiles these into local lattice surgery instructions that are allocated to hardware tiles according to a specified resource layout. The second stage retains logical parallelism while avoiding resource contention in the fault-tolerant layer, aiding realism. Additionally, users can specify dedicated tiles at which magic states are replenished, enabling resource costs from the logical computation to be considered independently from magic state distillation and storage. We demonstrate the applicability of our pipeline to large practical quantum circuits by providing resource estimates for the ground state estimation of molecules. Finally, we find that variable magic state consumption rates in real circuits can cause the resource costs of magic state storage to dominate unless production is varied to suit.

97 MATHEMATICS AND COMPUTING↗

MemFriend: Understanding Memory Performance with Spatial-Temporal Affinity

In HPC applications, memory access behavior is one of the main factors affecting performance. Improving an application’s memory access behavior involves optimizing data layout and/or restructuring code, and requires studying spatial-temporal data locality. Existing data locality analyses focus on single-location metrics and are restricted to evaluating temporal locality. We introduce spatial-temporal affinity metrics that quantify temporal access proximity, forward access correlation, and nearby access correlation between pairs of memory locations. We describe methods for distinguishing between potential vs. realized affinity and for reasoning about affinity at multiple resolutions (3D, 2D, 1D). Finally, we construct spatial-temporal affinity signatures that classify memory behavior and that be used to reason about changes in software (data relayout, code refactoring) or hardware (caching, prefetching). We describe methods for signature visualization, interpretation, and quantitative comparison of signatures. We evaluate our methodology using applications with variants that contrast data structures, data layouts and algorithms. We show that spatial-temporal affinity analysis provides novel insights and enables predictive reasoning about application performance when contrasted with reuse distance analysis.

Suriyakumar, Yasodhadevi↗

Exploration ToolKit (ExTK)

The Exploration Toolkit (ExTK) is a reusable Extended Reality (XR) system developed for incorporating and exploring 3-Dimensional (3-D) computer aided design (CAD) models in XR, with a primary focus on Augmented Reality (AR). The ExTK consists of a Developer Mode and a User Mode. In Developer Mode, ExTK provides developers with the ability to easily import 3-D CAD models and activate desired exploration functionality and layout. Multiple models can be added to a single instantiation of the ExTK using Unity's Scene capability. Exploration functions include scaling, rotating, explode/contract, animations, hiding parts, submodules, and measurement functions. In User Mode, ExTK provides a menu system that allows users to select models and initiate exploration functions. ExTK is architected for reusability and developers can customize the ExTK layout and functions according to application needs. ExTK is designed to be hardware agnostic, although initial development focused on the Microsoft HoloLens as the primary deployment platform. The ExTK is developed using the Unity Game Engine Development Platform. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525. SAND2021-1506 O

Klein, BrandonThorin↗

Experimental study of airpath electrification in an opposed-piston two stroke (OP2S) engine architecture

The opposed-piston two stroke (OP2S) engine shows potential as an alternative engine architecture to the conventional four stroke engine due to its high-power density, thermal efficiency, and versatile airpath management system. Since the pistons of a two-stroke engine do not pump the air into and out of the cylinder like in a four-stroke engine, the selection of the air induction devices and airpath actuators becomes critical to optimize engine performance. Both the pumping losses and the in-cylinder combustion process can be affected by the scavenging process in a two-stroke engine. Therefore, this study compares two different airpath configurations for the same family of OP2S engines and investigates performance metrics like scavenging control, pumping work, net indicated and brake efficiencies, and engine-out emissions associated with each airpath. Data was collected on a 3.2 L, two-cylinder OP2S engine with an electrically assisted turbocharger (EAT) and a 4.9 L displacement, three-cylinder engine with a variable geometry turbocharger (VGT) and a supercharger. The experiments consisted of speed and load sweeps for both engines at the same operating conditions to compare scavenge control in both architectures. For the three-cylinder layout, the SE sweep range was much higher, and the intake pressure could be independently varied with air flowrate, thus providing more flexibility for scavenging control. The supercharger and the VGT usage was optimized based on its efficiency map and thus, this layout had lower pumping losses compared to the EAT. The two-cylinder engine had a higher overall SE as compared to the three-cylinder engine, but the intake pressure and air flowrate could not be decoupled, leading to over scavenging and increased short circuiting of fresh charge into the exhaust.

Bhatt, Ankur [Clemson University, Clemson, SC, USA↗

Earthquake response of head-mounted equipment in advanced nuclear reactors

The seismic response of safety-related equipment mounted on the head of an advanced reactor, including pumps, control rod drive mechanisms, and reactor monitoring devices, will affect the design and layout of many advanced reactors. High earthquake-induced accelerations in such equipment may challenge their seismic qualification and trigger the need for additional support framing on the reactor head. Base isolation is a design solution that can drastically reduce seismic demands on equipment. This article describes a set of earthquake-simulator experiments conducted on a scale-model of a base-isolated reactor vessel including four representations of head-mounted equipment, with frequencies spanning from 4.5 to 27 Hz. Dynamic responses of the head-mounted equipment, including displacements, accelerations, and strains, were measured in the experiments for three support conditions: conventional, and seismically isolated using single concave Friction Pendulum (SFP) bearings and triple Friction Pendulum (TFP) bearings. Seismic isolation was effective at reducing equipment responses (accelerations, displacements, and strains) with respect to those in the conventionally supported vessel across a range of seismic inputs. Companion numerical studies highlight the accuracy to be expected in the calculation of different response quantities for lightly damped equipment. The importance of characterizing damping in head-mounted, safety-related equipment through physical experiments to support design and risk assessment is made clear through the numerical simulations.

Engineering↗

Two-stage formation-energy correction (NbZr, TaZr, VZr)

This bundle contains the scripts, the raw and corrected per-structure data, and the manuscript plots for the NbZr / TaZr / VZr BCC binary formation energies and the associated RMSDs. Why a two-stage correction is necessary: The "raw" formation energy of every relaxed VASP configuration is computed in the usual way, FE_raw(c) = E_alloy(c) - sum_i x_i * E_pure_i , where E_pure_i are the per-atom total energies of the pure-element reference structures (Nb, Ta, V, Zr in the same BCC supercell, with identical INCAR / KPOINTS / PAW choices). With perfectly consistent reference runs the raw FE should vanish at the two pure-element endpoints (x = 0 and x = 1) by construction. In practice this does not hold for two reasons that are present in our dataset: 1. Reference-energy inconsistency (composition-dependent bias). Even with identical input parameters, the pure-element runs (stored in `corrected_DFT_pure_element_runs/`) differ slightly from the values that would be implied by the alloy runs at near-pure compositions (a few meV/atom). This bias is approximately linear in concentration, because the residual error in E_pure_Nb (or E_pure_Ta / E_pure_V) propagates into FE_raw(c) as (1 - x) * dE_pure_1, and the corresponding error in E_pure_Zr propagates as x * dE_pure_2. Left uncorrected, this produces a non-physical "tilt" of FE_raw(x) and shifts the entire FE-vs-x cloud away from zero at the endpoints. 2. Endpoint anchoring against the audited true endpoints. The strict endpoint values (FE_x0_meVatom, FE_x1_meVatom in `corrected_fe_strict_endpoints_20260518/strict_endpoint_check_20260518.csv`) were re-derived from an independent cross-check of the pure-element runs. After stage 1 removes the linear bias, the near-pure compositions in the alloy dataset still extrapolate to values that differ slightly from these audited endpoints — because stage 1 is fit from a few near-end alloy bins, not from the audited pure-element references themselves. The README.txt file discusses how these issues are addressed by the two-stage correction, and describes folder layout, pipeline summary, and how to re-run.

36 MATERIALS SCIENCE↗

CONCEPT OF A POLARIZED POSITRON SOURCE FOR CEBAF

Positron beams would provide new and meaningful probes for the experimental program at the Thomas Jefferson National Accelerator Facility (JLab), including but not limited to future hadronic physics and dark matter experiments. Critical requirements involve generating positron beams with a high degree of spin polarization, sufficient intensity and a continuous-wave (CW) bunch train compatible with acceleration to 12 GeV at the Continuous Electron Beam Accelerator Facility (CEBAF). To address these requirements, a polarized positron injector based upon the bremsstrahlung of an intense CW spin polarized electron beam is considered*. First a polarized electron beam line provides >1 mA of polarized electrons at ~120 MeV to a high-power target for positron production. Next, a second beam line collects, shapes and aligns the spin of positrons for users. Finally, the positron beam is matched into the CEBAF acceptance for acceleration and transport to the end stations with energies up to 12 GeV. An optimized layout to provide positrons beams with intensity >100 nA (polarized) or intensity >3 µA (unpolarized) will be discussed in this poster.

Habet, S. H.↗

Positron beams at Ce+BAF

Positron beams would provide a new and meaningful probe for the experimental program at the Thomas Jefferson National Accelerator Facility (JLab). The JLab Positron Working Group, formed in 2018 and now with over 250 members from 75 institutions, continues to develop an experimental program with high duty-cycle positron beams including but not limited to future hadronic physics and dark matter experiments. Critical requirements involve generating positron beams with a high degree of spin polarization, sufficient intensity and a continuous-wave (CW) bunch train compatible with acceleration to 12 GeV at the Continuous Electron Beam Accelerator Facility (CEBAF). In this presentation we describe a start-to-end layout for positron beams at 12 GeV CEBAF utilizing the Low Energy Research Facility (LERF) at Jefferson Lab to build two new injectors. A GaAs dc high voltage photo-gun first generates >1 mA of polarized electrons which are then accelerated to 80-150 MeV and directed to a high-power spinning W target for polarized bremsstrahlung and positron pair creation. A second injector then collects, bunches and accelerates the positrons to 123 MeV. The positron beams are transported by a new beam line and injected into the CEBAF acceptance for acceleration to the end stations with energies up to 12 GeV. The layout is optimized to provide Users with positron spin polarization >60% and intensity greater than >100 nA, and with higher intensities when polarization is not required.

Benesch, J.↗

Task Force Report: ESR Linear Lattice Design

A new layout and optics for the Electron Storage Ring (ESR) have been produced with revised spin rotators in IRs 6 and 8, a redesign of IR10, and a new geometric layout for the ring. In this report, the efforts to produce this lattice are documented, including the motivation for design decisions. This new lattice, version 5.5, will be used as the baseline for future studies.

43 PARTICLE ACCELERATORS↗

Techno-Economic Wind Blade Manufacturing Model to Identify Opportunities for Cost Improvements Phase II IACMI Project 4.6/4.8

In IACMI Project 4.6 and IACMI Project 4.8, an Excel-based Techno-Economic Model (TEM) of the manufacturing process for composite wind turbine blades and a DELMIA Factory Flow Simulation of a generic wind blade manufacturing facility was developed. Together, these two tools provide a combined economic modeling capability that accounts for the material, labor, overhead and full-lifecycle operating costs associated with wind blade manufacturing as well as the impact of process flow and factory layout on overall manufacturing efficiency. The tools provide a novel means of detailed comparative analysis of the economic feasibility of proposed technologies and process changes for blade manufacturing. The modeling tools were developed with close support from members of industry and visits to multiple blade manufacturing facilities. With industry oversight, a detailed generalized manufacturing process plan and facility layout were developed with manufacturing parameters, material costs and economic factors based on historical data. Dassault Systèmes and the University of Texas at Dallas (UTD) contributed to the development of the Techno-Economic Model by providing macros to enable the generation of Bill of Material (BOM) data from a 3D blade design in either CATIA or NuMAD format, respectively. The TEM was built with the capability to directly import a Bill of Materials for economic analysis, and with the addition of the macros provided by Dassault and UTD, the TEM can directly import blade designs from both CATIA and NuMAD file formats. The modeling tools developed in Project 4.6 were used to investigate four wind blade manufacturing concepts in detail and select one to explore with laboratory-scale experimentation in Project 4.8. The four manufacturing concepts that were investigated were down-selected by the full project team from a larger list of concepts. The selections were made based on a number of criteria ranking viability and level of interest for each concept. The ‘One-Step Close’ manufacturing concept was ultimately selected for investigation in Project 4.8 and the demonstration was performed at the NREL CoMET facility. The TPI advanced manufacturing facility in Warren, RI contributed the production of several prototype components, the designs for which were developed by Janicki Industries. The demonstration project provided clear indication of the viability of the One-Step Close manufacturing concept for blade manufacturing and good validation of the Techno-Economic Model’s prediction of its economic impact.

17 WIND ENERGY↗

Advanced Ceramic Membranes/Modules for Ultra Efficient Hydrogen (H2) Production/Carbon Dioxide (CO2) Capture for Coal-Based Polygeneration Plants: Fabrication, Testing, and CFD Modeling

Inorganic membrane-based systems are a promising technology for precombustion CO2 capture with simultaneous H2 production. State-of-the-art packages for high temperature and pressure service consist of multiple tube membrane bundles prepared in a "candle filter" configuration, in which the membrane tubes are open at one end and sealed at the other. This configuration is used for practical reasons, specifically the need to minimize problems due to thermal expansion mismatch between the ceramic tube bundle and the steel housing. However, the primary technical problem with the candle filter format for commercial-scale installations is the inability to purge the tube side (typically the permeate side), a feature that is crucial for high H2 recovery. In this study, the focus is to design and fabricate the first dual-end open full ceramic multiple tube membrane bundle that enables tube side (permeate) purge for gas separation applications. An additional key feature of this design is the simplified module layout, as the membrane bundles can be installed end-to-end with tube side (permeate) flow directly from one bundle to the next. This layout simplifies the membrane to housing seals and yields significant improvement in membrane packing density. Detailed focus areas in our studies include: (i) Materials development and preparation of the tube-to-tube sheet potting for the dual-ended bundle; (ii) the sealing and optimal module configuration design to minimize membrane stress upon module mounting; (iii) the demonstration, via the fabrication of CMS and Pd-alloy membranes supported on full-size, dual-ended ceramic support bundles, of the first example of a purgeable ceramic membrane and module; and (iv) development of a CFD model of the membrane module for calculation of feed flow distribution, and for use in scale-up, and capital cost estimating. The CFD model was validated using experimental data with the multi-tubular membrane system, employing He/N2 as a model gas mixture (surrogate for H2/CO2), and has been shown to be quite accurate. Employing the model, we are able to study the effects of operating pressure and temperature, feed and sweep gas flow rates, and the choice of membrane tube configuration on system performance.

20 FOSSIL-FUELED POWER PLANTS↗

VOLTTRON Modular Framework: Enabling flexible and scalable deployment solutions

VOLTTRON ™ is an open-source platform for distributed sensing and control. The platform provides services for collecting and storing data from buildings and devices and provides an environment for developing applications which interact with that data. The platform allows developers to build out their use cases by utilizing these frameworks and integrating new capabilities. To simplify the deployment of systems built on VOLTTRON, a new way of organizing the codebase is being explored. This document details these efforts through a new code repository layout for the VOLTTRON platform and services, and how this new layout provides targeted deployments using standard python deployment packages (wheels). In addition, this paper will discuss the development of third-party agents and how they can be integrated within the VOLTTRON ecosystem. Finally, we will discuss core platform development and direction for the modularized version of VOLTTRON.

47 OTHER INSTRUMENTATION↗

Advanced Ceramic Membranes/Modules for Ultra Efficient Hydrogen (H2) Production/Carbon Dioxide (CO2) Capture for Coal-Based Polygeneration Plants: Fabrication, Testing and CFD Modeling

Inorganic membrane-based systems are a promising technology for precombustion CO2 capture with simultaneous H2 production. State-of-the-art packages for high temperature and pressure service consist of multiple tube membrane bundles prepared in a "candle filter" configuration, in which the membrane tubes are open at one end and sealed at the other. This configuration is used for practical reasons, specifically the need to minimize problems due to thermal expansion mismatch between the ceramic tube bundle and the steel housing. However, the primary technical problem with the candle filter format for commercial-scale installations is the inability to purge the tube side (typically the permeate side), a feature that is crucial for high H2 recovery. In this study, the focus is to design and fabricate the first dual-end open full ceramic multiple tube membrane bundle that enables tube side (permeate) purge for gas separation applications. An additional key feature of this design is the simplified module layout, as the membrane bundles can be installed end-to-end with tube side (permeate) flow directly from one bundle to the next. This layout simplifies the membrane to housing seals and yields significant improvement in membrane packing density. Detailed focus areas in our studies include: (i) Materials development and preparation of the tube-to-tube sheet potting for the dual-ended bundle; (ii) the sealing and optimal module configuration design to minimize membrane stress upon module mounting; (iii) the demonstration, via the fabrication of CMS and Pd-alloy membranes supported on full-size, dual-ended ceramic support bundles, of the first example of a purgeable ceramic membrane and module; and (iv) development of a CFD model of the membrane module for calculation of feed flow distribution, and for use in scale-up, and capital cost estimating. The CFD model was validated using experimental data with the multi-tubular membrane system, employing He/N2 as a model gas mixture (surrogate for H2/CO2), and has been shown to be quite accurate. Employing the model, we are able to study the effects of operating pressure and temperature, feed and sweep gas flow rates, and the choice of membrane tube configuration on system performance.

20 FOSSIL-FUELED POWER PLANTS↗

Data Structure Alchemy

In an increasingly more data-driven world, the project set out to uncover the first principles of data-structure design, chart the immense design space they form, and build automation that can synthesize an optimal structure, or even a whole storage engine, for any given workload, hardware platform, and cost target. Data structures are at the center of every computational system and are directly responsible for its performance. Two core technical thrusts were defined: 1) Mapping design spaces for key data-centric abstractions (filters, hash functions, storage-engine layouts, neural-network topologies, blockchain protocols, image layouts, etc.). 2) Developing search & synthesis algorithms, initially analytical cost models, later neural-guided bi-level optimisers that navigate sextillions of candidate designs in seconds and materialise the best one as ready‐to-run code. This report distills the key insights, accomplishments, and impact.

97 MATHEMATICS AND COMPUTING↗

Organic Direct-Bonded-Copper-Based Rapid Prototyping for Silicon Carbide Power Module Packaging

Silicon carbide (SiC) power devices are playing ever- growing roles in high-power-density power electronics converters by offering benefits such as high voltage rating, fast transients, and high thermal performance. Organic direct-bonded copper (ODBC)-based packaging, due to its ductility and ease of han- dling, allows the possibility of a more flexible layout design that may better tap the potential of SiC benefits. In this work, an ODBC-based prototyping routine is developed that accelerates the iterations of packaging layout design with low cost. The properties of ODBC and its handling are briefly introduced, and tools and fabrication steps are explained. Following this routine, a 1.2-kV SiC half-bridge power module is designed and fabricated with the focus on sub-nanohenry ultra-low loop inductance. Simulation and experimental validation are also conducted.

25 ENERGY STORAGE↗

Flexible Fuel Electric Hybrid Glass Furnace Demonstration

The furnace contractor completed General Arrangement (GA) layouts and Piping and Instrumentation Diagrams (P&ID) for both furnaces during this reporting period. Preliminary equipment lists accompanied this work. Control system architecture in progress. Key furnace design elements and deliverables on track to be complete by end of phase. The facilities design and integration contractor completed a 3D scan and modeling of the facility. The work during this period focused on preliminary design with a focus on overall project scope and feasibility. This key work supports the overall project deliverables for 30% preliminary engineering and Preliminary Design Report (PDR) which are on track to be complete by the end of the phase. These deliverables are approximately 50% complete through the end of this reporting period. The batch house operations design contractor continues development of site plan and General Arrangement (GA) layout to support the 30% preliminary engineering and design package. This work is 60% complete by the end of this reporting period and is on track to be complete by the end of the phase.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Spectrum Unfolding with the MC-15

The Multiplicity Counter 15 tube detector or MC-15 is an optimized detector designed for use in the field. It is composed of 15 3 He tubes embedded in high density polyethylene (HDPE). Recent work has explored expanding the use of the MC-15 beyond multiplicity counting to neutron dosimetry applications. Knowledge of the neutron energy spectrum information is required to use a detector as a neutron dosimeter. The MC-15 tube layout is shown in Figure 1. The unique layout makes it possible to use the detector for neutron spectroscopy via spectrum unfolding. Spectrum unfolding requires (1) energy dependence of the detector response, (2) a detector response matrix that precisely quantifies the response to mono-energetic neutrons, (3) an initial guess spectrum, (4) an unfolding algorithm, and (5) measured data (counts in the case of the MC-15). An energy dependent detector response matrix (DRM) can be constructed by considering either each of the three rows of 3 He tubes as a distinct detector or each individual tube as a distinct detector. The HDPE separating the 3 He in the MC-15 provides the distinct energy dependent response for the rows and individual tubes. In this report we detail the development of detector response matrices for the MC-15 and the application of the Los Alamos Unfolding Code (LUC) to both simulated and measured data. Three MC-15 orientations were studied: (1) standard orientation with the MC-15 front facing the source, (2) standard orientation with Cd sheet, (3) 90° orientation with the side of the MC-15 facing the source.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗