Engineering PapersSearch

SEARCH · Engineering Papers

Results for “Architecture”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Grid Architecture Mapping to Understand Transformation (GAMUT): Methods and Framework Architecture

Grid architecture (GA) is a concept that was developed to address the need for a comprehensive view of power grid challenges. GA can be viewed as a relatively consistent and fixed high-level approach; however, for any instantiation of grid structures, a combinatorial explosion results from each lower-layer expansion. This constitutes the main challenge with GA—it is a grid architect’s view of the system, which might not be very informative at the implementation level. Grid Architecture Mapping to Understand Transformation (GAMUT project) seeks to bridge that gap by integrating subject matter expertise across GA structures, providing users who lack expertise in GA approaches with valuable insights and informational materials. GAMUT seeks to answer feasibility questions for the approach. System-level expectations are that a GA baseline needs to be established in order for GA to be the common framework to which any lower layer approach is tied. This report explores a potential information ingestion and documentation framework to support GAMUT. The main concepts that enable the solution domain of GAMUT are discussed, and examples are provided. The solution domain leverages already-existing technology and concepts related to GA, knowledge management, and other relevant areas. To assess GAMUT building blocks and the overall approach, a feasibility assessment is proposed, rooted in systems engineering and GA architecture evaluation concepts.

24 POWER TRANSMISSION AND DISTRIBUTION

Solid‐State Prealkylation of Electrode Architectures (SPEAR): Direct Control of Prelithiation Levels in Silicon Anodes and Electrochemical Cycling

The Solid-state Prealkylation of Electrode ARchitectures (SPEAR) is different than traditional electrochemical prealkylation processes. Through SPEAR, alkylation is driven by solid-state diffusion without the simultaneous SEI formation concomitant with polarization. Here, we investigate the prelithiation of 80 wt. % Si-based anodes to varying amounts (up to Li 1.38 Si) to understand the trade-off between improved Li capacity and expansion-induced stress. Through dilatometry, we found that solid-state lithiation led to filling of the electrode pores through silicon expansion. This swelling changed the SEI formation process and accessibility of the silicon compared to an electrochemically lithiated electrode. Indeed, optimal prelithiation to Li 0.82 Si increases the initial C/3 cycling capacity post-SEI formation up to 43%, consistent with deeper Si activation through the electrode bulk. Prelithiation and cycling cells prelithiated beyond Li 0.82 Si results in a state of charge (SOC) close to 100% which facilitates parasitic degradation mechanisms and volume expansion of the Si electrode. The results demonstrate a pathway to modify silicon activation/SEI formation to enable high-energy electrodes.

Musgrove, Amanda L. [Oak Ridge National Laboratory

Comparison of Real-Time Pressure Rail Selection Algorithms for the Hybrid Hydraulic Electric Architecture: Case Study on a Track Loader

Abstract The hybrid hydraulic electric architecture (HHEA) seeks to combine the high power/torque/force density of hydraulics with the efficiency of electric machines. A set of common pressure rails is used to provide a majority of the power and this power is modulated by small electric machines to provide precise control for the operator. The HHEA has been studied in previous work using off-line dynamic programming optimization to determine energy efficient pressure rail selections, but this approach requires drive cycle information apriori. A Lagrange multiplier method has also been investigated where a set of gains (Lagrange multipliers) are optimized off-line with the idea the these gains, once determined, could be used for real-time operation. In this work, three new real-time pressure rail selection algorithms that do not require future drive cycle information are investigated; greedy, torque minimizing, and thresholding. The greedy control is found to only use 1% more energy than the globally optimal dynamic programming solution; but a model of energy loss is required.

24 POWER TRANSMISSION AND DISTRIBUTION

Pulsed Infrared Thermography Nondestructive Imaging of SiC-SiC f Composite Cladding Architectures; Understanding the Performance of SiC-SiC f Composite Cladding Architectures with Cr Coating in Normal Operating and Accident Conditions in LWRs and Advanced Reactors

SiC-SiC f composites, consisting of silicon carbide fibers embedded in a silicon carbide matrix, are advanced materials with high thermal conductivity, temperature stability, and resistance to radiation damage. Traditional methods for quality control of fabricated SiC-SiC f composites involve nondestructive evaluation (NDE) with X-ray computed tomography (XCT). However, XCT imaging of typical SiC-SiC f structures for cladding applications can involve several hours. In this project, we investigate an alternative approach to NDE of SiC-SiC f composites that involves rapid (on the order of seconds) imaging with Pulsed infrared thermography (PIT). PIT images of a planar SiC-SiC f specimen show the structure of the surface monolith layer and internal SiC f structures. The capability of PIT imaging in visualizing SiC f structures is qualitatively confirmed by observing similarity in the PIT and X-ray transmission images of the same specimen. Computer vision analysis of defects in the PIT image of the monolith was performed with thresholding followed by topological structural analysis that computed geometric descriptors, including major/minor axes of fitted ellipses, area, perimeter, Feret diameter, circularity, roundness, and solidity.

22 GENERAL STUDIES OF NUCLEAR REACTORS

Open Architecture for Cost Savings in Advanced Nuclear Reactors

Recently, nuclear power plant build projects in the West have run over budget due to high capital costs and schedule overruns. Compared to other sources of energy, nuclear power plants have higher capital costs. Reactors are often different at every site, resulting in a lack of standardization. Nuclear is expected to compete with other low carbon sources of energy which have lower capital costs making it essential for nuclear to develop ways of reducing costs. Strategies such as standardization, learning rates, modularization, and schedule reduction in advanced reactors can reduce nuclear costs by about 40%. Standardization as a way of cutting capital costs has been explored even in large nuclear power plants. Standardization of certain plant components can result in lower component and installation costs and higher learning from experience. Standardization can be achieved by adopting a criterion of key performance indicators and general design principles for a specific system or component such as the balance of plant. Modularization allows the construction of certain components of SMRs in a factory, which saves time, increases productivity, and encourages higher learning rates. Production learning decreases the time and the cost related to an activity. The potential for modularized components of advanced reactors to be manufactured in factories makes it conducive to achieving higher learning rates. Developing large-capacity nuclear programs through sequential builds cultivates a higher learning rate, which in effect may reduce schedule overruns. Open architecture has been identified as a way to drive standardization among advanced reactor designs and result in cost savings. Open architecture (OA) is defined as a design enabling a diverse supply chain by defining and publishing requirements of systems or equipment in functional and/or interface terms, utilizing technical standards in widespread use. Currently, the nuclear industry’s approach is to use closed architecture, making most designs proprietary. However, collaboration between various advanced reactor vendors and suppliers utilizing the concept of open architecture can result in modular and standardized architecture of subsystems or subcomponents of a nuclear power plant. Completely standardizing nuclear power plants may be impossible, however, certain common subsystems amongst the various reactor designs could be standardized and/or access a wider supply chain and leverage existing learning from other sectors. Open architecture will save time and allocate resources to the parts of the plants that have the most unique features. A key advantage of open architecture is its ability to improve production learning across advanced reactors (AR) types in the industry, by providing and utilizing the same kind of component. Sodium fast reactor (SFR), High Temperature Gas Reactor (HTGR) and Molten Salt Reactor (MSR) are the advanced reactors considered for this project. This paper aims to determine the cost savings in advanced reactor programs due to open architecture learning rate. This work is an extension of work done on light water reactor small modular reactors; the cost methodology was utilized to investigate the impact of open architecture on advanced reactors with a particular focus on sodium fast reactors. The cost data on sodium fast reactors used in the model presented the most adequate information required for the analysis.

Advanced Nuclear Reactors

Fundamental Studies of the Vibrational, Electronic, and Photophysical Properties of Tetrapyrrolic Architectures

The ability to capture and utilize light in the near-ultraviolet (NUV), visible and near-infrared (NIR-I and NIR-II) spectral regions (i.e., 320–400, 400–700, 700–1000, 1000–1700 nm) is essential for any solar-energy conversion scheme. Nature employs chlorophylls and bacteriochlorophylls in light-harvesting architectures to absorb light in the blue and red/NIR regions. Accessory pigments (carotenoids, bilins) augment absorption of the (bacterio)chlorophylls in the green region. The harvested energy is funneled to a reaction center protein, where charge separation occurs. Subsequent migration of the electron and the hole stabilizes and stores the energy from light via redox chemistry. The long-term objective of the Bocian/Holten&Kirmaier/Lindsey research program under this DOE grant has been to develop tetrapyrrole-based molecular architectures that absorb sunlight, funnel energy and separate charge with high efficiency. Integral to the program has been iterative cycles of design, synthesis and characterization that provided deep insights into the relationships between chemical composition, electronic structure, and key static and dynamic properties (vibrational, redox, photophysical, energy/charge transfer) of tetrapyrrolic systems. Such architectures included monomers, dyads, larger arrays, and complexes with accessory components. The objective was to develop molecular designs and guiding principles to enhance current and future energy-conversion schemes. Molecular arrays targeted to address one or more fundamental questions concerning light harvesting and energy/charge transfer were constructed from analogues of the naturally occurring hemes, chlorophylls and bacteriochlorophylls. Diverse, tunable synthetic building blocks were prepared that spanned the three respective tetrapyrrole families, which are the porphyrins, chlorins and bacteriochlorins. Thus, the research focused on porphyrins as well as synthetic surrogates for chlorophylls (chlorins, 13 1 -oxophorbines and chlorin-imides) and bacteriochlorophylls (bacteriochlorins, bacterio-13 1 -oxophorbines and bacteriochlorin-imides), generically termed hydroporphyrins. Although the three tetrapyrrole classes (porphyrins, chlorins and bacteriochlorins) absorb light strongly in the violet-blue spectral region, the long-wavelength absorption band typically lies in the green-orange, red, and NIR regions, respectively, with increasing intensity. Understanding the spectra, electronic structure, and energy/charge-transfer properties of such tetrapyrrolic macrocycles is of central importance for the rational design of molecular architectures for solar-energy conversion. Our integrated program of molecular design and synthesis coupled with a variety of spectroscopic, electrochemical, and computational studies have probed from first principles how structural and electronic properties of tetrapyrrolic macrocycles dictate spectral properties as well as the rates of ground-state hole/electron transfer and excited-state energy flow in multicomponent architectures. Individual molecules and multicomponent architectures were designed to test ideas of fundamental importance, often requiring the development of new synthetic methodology. The members of the collaborative team had almost daily discussions by phone and/or e-mail concerning design of molecules, flow of compounds between the labs, planning of physical characterization studies, discussing results and analysis and integrating into design of next generation architectures, and the preparation of manuscripts. Furthermore, students and postdocs in the different labs routinely communicated with one another to facilitate the advancement of the research activities. In short, a highly integrated and collaborative research program was well established among the groups. The research effort involved molecular design and synthesis of synthetic molecular architectures by the Lindsey group integrated with physicochemical and photophysical characterization by the Bocian group and the Holten&Kirmaier group (Figure 2). The Bocian group carried out electrochemical, electron paramagnetic resonance (EPR), resonance Raman (RR), and Fourier-transform infrared (FT-IR) studies, as well as density functional theory (DFT) calculations and the time-dependent extension (TDDFT) to gain insight into excited-state properties. The Holten&Kirmaier group carried out static and time-resolved absorption and fluorescence spectroscopy studies and simulated absorption spectra using molecular orbital (MO) energies from DFT as input to the four-orbital model to complement the TDDFT calculations. The combined measurements provided understanding of the vibrational/electronic properties of the individual molecules and the changes that occur upon incorporation into multicomponent architectures. This information underpinned elucidating the mechanisms and timescales of ground-state hole/electron transfer and excited-state energy and charge transfer.

14 SOLAR ENERGY

HARMONY: Large-Scale Architecture Search for Efficient Hybrid Language Models

As large language models scale to trillions of parameters, their computational and memory requirements present critical challenges for efficient training and deployment. While Mixture of Experts (MoE) architectures enable efficient scaling through sparse parameter activation, and state-space models like Mamba offer linear-time complexity, principled methods for combining these paradigms remain undeveloped. We introduce HARMONY (Hybrid Architecture Research for Mamba, Optimized with Neural efficiencY), a multi-objective evolutionary neural architecture search framework for discovering efficient hybrid language models that integrate Transformer attention mechanisms, Mixture-of-Experts routing, and Mamba state-space components. Through large-scale distributed search using 16,384 MI250X GPUs on the Frontier supercomputer, HARMONY explores a comprehensive design space encompassing six attention variants (MHA, MQA, GQA, MLA, SWA, and Mamba-2), variable MoE configurations with both routed and shared experts, and extensive Mamba hyperparameters. Our framework discovers heterogeneous architectures that balance training performance with computational efficiency through multi-objective optimization incorporating latency penalties and fitness-based selection. Analysis of discovered architectures reveals that optimal hybrid designs favor heterogeneous component mixing rather than homogeneous patterns, with Mamba-2 and Multi-Head Latent Attention (MLA) emerging as preferred mechanisms. Discovered architectures demonstrate superior training efficiency: our best configuration achieves a final perplexity of 1.0874 with 2.38B parameters while processing 4,320 tokens/second, outperforming significantly larger manually designed models. Full-scale evaluation shows HARMONY's top architectures achieve better loss trajectories than equivalently-sized models using state-of-the-art configurations including Mixtral, Jamba, and Samba. Additionally, we demonstrate 91% weak scaling efficiency when training discovered 36B-parameter models across 1,024 GPUs. HARMONY is released as an open framework with comprehensive tools for building and training hybrid models using expert-data-pipeline parallelism, democratizing access to automated architecture design for next-generation language models.

Herron, Emily [ORNL] (ORCID:0000000273008172)

Land-Based Wind Reference Architecture

This report describes a summary and the final deliverables for the Wind Reference Architecture project funded by the Department of Energy's (DOE) Wind Energy Technology Office (WETO). The project objective was to further refine the reference architecture of the existing wind power plant reference architecture developed by Idaho National Laboratory (INL) and Sandia National Laboratories (SNL) by expanding the wind turbine generator (WTG) into a separate individual wind turbine reference architecture. Additional details regarding the wind turbine system and how it integrates with a wind power plant were researched and reflected in the reference architecture. This addition is important because it allows researchers to understand the components and devices in a wind turbine, how they function and how a wind turbine integrates into a wind power plant to be able to perform cybersecurity evaluations on the system. The communication and control systems were the focus while developing this reference architecture to understand the cybersecurity posture of a wind turbine and power plant. The information technology (IT) and operational technology (OT) systems perspective are integral in helping improve the cybersecurity posture by creating a starting point to study how wind energy can be safely designed and deployed in our power grid from an integrated systems standpoint. A holistic wind power plant and turbine reference architecture and simulation was developed and made open-sourced for further research efforts.

17 WIND ENERGY

Approximate 𝑡-Designs in Generic Circuit Architectures

Unitary 𝑡-designs are distributions on the unitary group whose first 𝑡 moments appear maximally random. Previous work has established several upper bounds on the depths at which certain specific random quantum circuit ensembles approximate 𝑡-designs. Here we show that these bounds can be extended to any fixed architecture of Haar-random two-site gates. This is accomplished by relating the spectral gaps of such architectures to those of one-dimensional brickwork architectures. Our bound depends on the details of the architecture only via the typical number of layers needed for a block of the circuit to form a connected graph over the sites. When this quantity is bounded, the circuit forms an approximate 𝑡-design in at most linear depth. We give numerical evidence for a stronger bound that depends only on the number of connected blocks into which the architecture can be divided. We also give an implicit bound for nondeterministic architectures in terms of properties of the corresponding distribution over fixed architectures.

information scrambling

The Design and Evaluation of Zero Trust Architecture for Electric Vehicle Charging Infrastructure: EVs @ Scale Series on EV Charging Station Cybersecurity

Implementing a zero trust architecture can significantly bolster the security of electric vehicle (EV) charging infrastructure. EV charging infrastructure includes numerous networked interfaces, each of which can present potential vulnerabilities. When these vulnerabilities are exploited, they can compromise the entire system, leading to severe operational and security risks. Zero trust is a security model that operates on the principle of "never trust, always verify," which helps manage the attack surface and limit the scope of any potential compromises. Fundamentally, this model ensures that no entity, whether inside or outside the network, is trusted by default. The design principles of zero trust include continuous verification, strict deny-by-default access controls, and micro-segmentation. Continuous verification ensures that every request is thoroughly checked, regardless of its origin. Strict access controls enforce the principle of least privilege, allowing users and devices only the minimum necessary access to perform their functions. Micro-segmentation involves dividing the network into smaller, isolated segments to prevent lateral movement in case of a breach. In the context of EV charging infrastructure, zero trust can be implemented through various strategies. For example, multi-factor authentication (MFA) can be required for engineers to access the management interfaces and control systems of charging stations. Real-time monitoring and analysis of network traffic can help detect and respond to anomalies. Systems that do not need to communicate with each other can be micro-segmented to enhance security. All communications should adhere to predefined policies to be permitted. Additionally, encrypting communications can protect sensitive information exchanged between chargers and management systems. This paper presents a zero trust architecture specifically designed for EV charging infrastructure. Implementing zero trust not only mitigates risks but also builds a resilient infrastructure capable of withstanding and quickly recovering from cyber threats. The architecture addresses six defined security objectives. A comprehensive test plan is developed to assess the architecture against these objectives, and the results of the evaluation are reported. This approach is essential for maintaining the reliability and integrity of EV charging services in an increasingly interconnected and vulnerable digital landscape. This is the first in a planned series of papers exploring the implementation of zero trust in EV charging infrastructure. Each paper will delve into different aspects and applications of zero trust, highlighting how various work processes and requirements can lead to distinct architectural designs. These architectures will be tailored to address specific security challenges and operational needs within the EV charging ecosystem, ensuring a robust and adaptable security framework.

33 ADVANCED PROPULSION SYSTEMS

The Concept and Role of Reference Architectures In NIF LRU Refurbishment Factories within LLNL

The National Ignition Facility (NIF) at Lawrence Livermore National Laboratory (LLNL) operates one of the most advanced laser systems in the world, relying on a vast number of optical components and Line Replaceable Units (LRUs) to maintain its functionality. Over time, these components degrade due to operational wear, necessitating refurbishment to sustain performance. However, many NIF LRU refurbishment factories have been “mothballed” or suffer from aging infrastructure, inconsistent work flows, and inefficiencies due to different approaches to production control and management. This paper explores the concept of reference architecture as a standardized framework to guide the redevelopment and restructuring of NIF LRU refurbishment factories. By establishing a common reference architecture, the refurbishment process can achieve reduced inefficiencies, produce quality products, and enhanced coordination across factories. This paper evaluates existing reference architectures, particularly those that integrate technical architecture, business architecture, customer context perspectives, and proposes tailored reference architecture for NIF LRU refurbishment factories.

42 ENGINEERING

Two Artificial Leaf Architectures for Solar Formate Production From CO 2 and H 2 O

Sunlight-powered artificial leaves for the production of formate from CO 2 are an attractive route to solar fuels, yet existing solar formate devices remain low in performance, and their architecture and material choices are underexplored. Herein, we report the fabrication of two distinct fully integrated, self-standing solar formate device architectures and elucidate the underlying design principles and material selection strategies. The first architecture integrates a Si photocathode with a BiVO 4 photoanode and utilizes a highly active Pd catalyst for CO 2 reduction. It represents the first artificial leaf device comprising two photoelectrodes (excluding photovoltaic [PV]-biased electrodes) for effective formate production under single-beam illumination. The second architecture employs a dark cathode and a dark anode driven by a 4-junction perovskite solar cell and uses a highly stable Bi catalyst for CO 2 reduction. This device delivers a record-high formate production rate of 174 µmol h −1 with a remarkable solar-to-formate energy efficiency of 2% among all artificial leaf devices reported to date. Finally, these results demonstrate the feasibility and outline the design principles of both PV-free and PV-assisted device architectures in solar fuel production.

electrocatalysis

Holistic energy analysis method for thermal management architectures of data centers

Modern high-performance computing (HPC) data centers (DCs), particularly those supporting energy-intensive artificial intelligence (AI) workloads, face escalating thermal management challenges that degrade performance through thermal throttling and drive up cooling power consumption and operational costs. To address this challenge, many have developed a wide variety of thermal management solutions (single-phase, two-phase, direct, indirect, hybrid, and more) which attempt to cool HPC DCs effectively while attempting to minimize overall system power consumption. However, the analysis of these solutions and methods to effectively compare one with another is lacking. Overall power usage effectiveness (PUE) and total-power usage effectiveness (TUE) provide a metric to quantify power consumption but fail to identify components in the system which require further optimization. To address this, we propose a holistic analytical framework – the waterfall diagram (WFD) – which leverages a waterfall chart methodology, offering a comprehensive visualization of both the thermal management system loop and heat flow pathways from individual server components to the outdoor ambient. Use of the WFD enables graphical estimations of power efficiency and cooling performance across each component of a DC cooling system and complements Sankey-style energy flow visualizations by additionally resolving stage-wise temperature changes and incremental TUE contributions. The framework is used in conjunction with simulation-based approaches, to conduct a detailed pressure drop and flow distribution analysis aimed at identifying the optimal coolant distribution architecture for a single-phase direct-to-chip water-cooled DC, which serves as the baseline for subsequent WFD analysis. Among the evaluated architectures, the 3 U modular coolant distribution architecture is found to demonstrate the best performance, considering minimal pressure drop and uniform flow distribution. In addition, TUE is calculated for each cooling loop component based on its associated pressure drop and corresponding pumping power, which are integrated into the WFD. This correlation between TUE and local temperature offers immediate insight into the power efficiency and thermal performance contributions of individual components, facilitating further development and optimization. Examples of WFD applications are presented under varying thermal loads and ambient conditions, demonstrating reasonable cooling strategies. Notably, the 3 U modular architecture maintains a consistent chip case temperature of 85°C, achieving a TUE of 1.016 at ambient temperature of 47°C, and a TUE of 1.026 at ambient temperature of 52°C. The WFD methodology provides an efficient, holistic, and streamlined framework for DC thermal management architecture assessment and enables design optimization which is important for addressing the thermal-fluidic energy challenges of current and next-generation DCs.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI

UltraLiM: In-Memory Boolean Logic Architecture Using UltraRAM

Conventional computing architectures encounter ‘von Neumann’ and ‘memory wall’ bottlenecks which arise due to the back-and-forth data movement between the physically separate memory and processing units and the speed mismatch between them, respectively. These bottlenecks hurt both energy efficiency and the throughput of computing systems. To address these challenges, in-memory computing architectures have emerged as a promising alternative. They reduce the need for frequent data movement by executing different computing tasks inside the memory system. Here, we present UltraLiM, a logic-in-memory architecture using the UltraRAM-based memory system. UltraRAM holds the promise of developing a ‘universal memory’, overcoming the limitations of charge-based memories thanks to their non-volatile behavior with lower operating voltage. This work presents an in-memory computing architecture that integrates an UltraRAM-based memory array with a custom-designed peripheral circuitry. With this architecture, we can perform various in-memory Boolean logic operations (such as NOT, NAND, NOR, and XOR) in a single cycle. Leveraging the separate read-write paths in the UltraRAM-based memory array, we optimize read operations without encountering design conflicts. This optimization enhances the sense margin, enabling the use of simpler peripheral circuitry for in-memory logic operations.

Alam, Shamiul [University of Tennessee, Knoxville

From Edge to HPC: Investigating Cross-Facility Data Streaming Architectures

In this paper, we investigate three cross-facility data streaming architectures, Direct Streaming (DTS), Proxied Streaming (PRS), and Managed Service Streaming (MSS). We examine their architectural variations in data flow paths and deployment feasibility, and detail their implementation using the Data Streaming to HPC (DS2HPC) architectural framework and the SciStream memory-to-memory streaming toolkit on the production-grade Advanced Computing Ecosystem (ACE) infrastructure at Oak Ridge Leadership Computing Facility (OLCF). We present a workflow-specific evaluation of these architectures using three synthetic workloads derived from the streaming characteristics of scientific workflows. Through simulated experiments, we measure streaming throughput, round-trip time, and overhead under work sharing, work sharing with feedback, and broadcast and gather messaging patterns commonly found in AI-HPC communication motifs. Our study shows that DTS offers a minimal-hop path, resulting in higher throughput and lower latency, whereas MSS provides greater deployment feasibility and scalability across multiple users but incurs significant overhead. PRS lies in between, offering a scalable architecture whose performance matches DTS in most cases.

George, Anjus [ORNL] (ORCID:0000000179737061)

On a Simplified Approach to Achieve Parallel Performance and Portability Across CPU and GPU Architectures

This paper presents software advances to easily exploit computer architectures consisting of a multi-core CPU and CPU+GPU to accelerate diverse types of high-performance computing (HPC) applications using a single code implementation. The paper describes and demonstrates the performance of the open-source C++ matrix and array (MATAR) library that uniquely offers: (1) a straightforward syntax for programming productivity, (2) usable data structures for data-oriented programming (DOP) for performance, and (3) a simple interface to the open-source C++ Kokkos library for portability and memory management across CPUs and GPUs. The portability across architectures with a single code implementation is achieved by automatically switching between diverse fine-grained parallelism backends (e.g., CUDA, HIP, OpenMP, pthreads, etc.) at compile time. The MATAR library solves many longstanding challenges associated with easily writing software that can run in parallel on any computer architecture. This work benefits projects seeking to write new C++ codes while also addressing the challenges of quickly making existing Fortran codes performant and portable over modern computer architectures with minimal syntactical changes from Fortran to C++. We demonstrate the feasibility of readily writing new C++ codes and modernizing existing codes with MATAR to be performant, parallel, and portable across diverse computer architectures.

97 MATHEMATICS AND COMPUTING

A Study on the Impact of Temperature-Dependent Ferroelectric Switching Behavior in 3D Memory Architecture

The flourishing development of neural networks that require exponentially growing amounts of data has presented an elevated demand for memory footprint. To address this, researchers have been exploring hardware accelerators with innovative memory architectures like 3D memory. These 3D memory architectures offer enhanced storage capacity and processing capabilities, at a cost of rising on-chip temperature during operation. Hafnium Zirconium Oxide (HZO) based Ferroelectric Random Access Memory (FeRAM) is a promising nonvolatile memory candidate in neural network hardware accelerators for its outstanding write performance and reliability. However, its implementation in the architecture regarding the temperature-dependent ferroelectric switching behavior has not been well studied. In this work, we study the thermal impacts on polarization switching through experimental devices and simulation results. We conduct the circuit and architecture-level simulations to showcase that one can exploit this temperature rise to reduce FeRAM's write voltage and write energy due to its unique temperature-activated polarization switching mechanisms. As the on-chip temperature increases to 351K (ambient temperature at 300K) due to neural network workloads, the access energy per bit can be reduced by 27.6% when a dynamic write voltage is applied.

36 MATERIALS SCIENCE

Distinct Melt Infusion Architectures of Antiperovskite Solid Electrolytes

Antiperovskite solid electrolytes are an emerging class of lithium‐ion conductors distinguished by their unusually low melting points, enabling scalable, low‐temperature processing routes not accessible to most solid electrolytes. In this work, we synthesize phase‐pure chloride (Cl), bromide (Br), and mixed halide (ClBr) variants of antiperovskites and investigate their ionic conductivities in both powder and hot‐pressed forms. Hot pressing significantly enhances conductivity across all compositions, while energy‐dispersive X‐ray spectroscopy (EDS) of the mixed halide system reveals halide surface migration during densification. We further investigate the melt‐infiltration behavior of these electrolytes into substrates relevant to solid‐state battery architectures, including Al and Cu current collectors, conventional NMC and LFP cathodes, and a foamed NMC cathode with a highly porous architecture. The foamed cathode enables deep and uniform electrolyte penetration, highlighting the role of electrode architecture in facilitating melt infiltration. Across all substrates, electrolyte halide chemistry strongly influences wetting behavior, penetration depth, and resulting microstructural morphology. Together, these results establish clear processing–structure relationships for melt‐infiltrated antiperovskite solid electrolytes and demonstrate how electrolyte chemistry and electrode architecture govern interfacial morphology during integration, providing practical guidelines for processing and structural design in solid‐state battery systems.

antiperovskite