Engineering PapersSearch

SEARCH · Engineering Papers

Results for “Architecture”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

A Study on the Impact of Temperature-Dependent Ferroelectric Switching Behavior in 3D Memory Architecture

The flourishing development of neural networks that require exponentially growing amounts of data has presented an elevated demand for memory footprint. To address this, researchers have been exploring hardware accelerators with innovative memory architectures like 3D memory. These 3D memory architectures offer enhanced storage capacity and processing capabilities, at a cost of rising on-chip temperature during operation. Hafnium Zirconium Oxide (HZO) based Ferroelectric Random Access Memory (FeRAM) is a promising nonvolatile memory candidate in neural network hardware accelerators for its outstanding write performance and reliability. However, its implementation in the architecture regarding the temperature-dependent ferroelectric switching behavior has not been well studied. In this work, we study the thermal impacts on polarization switching through experimental devices and simulation results. We conduct the circuit and architecture-level simulations to showcase that one can exploit this temperature rise to reduce FeRAM's write voltage and write energy due to its unique temperature-activated polarization switching mechanisms. As the on-chip temperature increases to 351K (ambient temperature at 300K) due to neural network workloads, the access energy per bit can be reduced by 27.6% when a dynamic write voltage is applied.

36 MATERIALS SCIENCE

Distinct Melt Infusion Architectures of Antiperovskite Solid Electrolytes

Antiperovskite solid electrolytes are an emerging class of lithium‐ion conductors distinguished by their unusually low melting points, enabling scalable, low‐temperature processing routes not accessible to most solid electrolytes. In this work, we synthesize phase‐pure chloride (Cl), bromide (Br), and mixed halide (ClBr) variants of antiperovskites and investigate their ionic conductivities in both powder and hot‐pressed forms. Hot pressing significantly enhances conductivity across all compositions, while energy‐dispersive X‐ray spectroscopy (EDS) of the mixed halide system reveals halide surface migration during densification. We further investigate the melt‐infiltration behavior of these electrolytes into substrates relevant to solid‐state battery architectures, including Al and Cu current collectors, conventional NMC and LFP cathodes, and a foamed NMC cathode with a highly porous architecture. The foamed cathode enables deep and uniform electrolyte penetration, highlighting the role of electrode architecture in facilitating melt infiltration. Across all substrates, electrolyte halide chemistry strongly influences wetting behavior, penetration depth, and resulting microstructural morphology. Together, these results establish clear processing–structure relationships for melt‐infiltrated antiperovskite solid electrolytes and demonstrate how electrolyte chemistry and electrode architecture govern interfacial morphology during integration, providing practical guidelines for processing and structural design in solid‐state battery systems.

antiperovskite

Unveiling shared genetic regulators of plant architectural and biomass yield traits in the Sorghum Association Panel

Abstract Sorghum is emerging as an ideal genetic model for designing high-biomass bioenergy crops. Biomass yield, a complex trait influenced by various plant architectural characteristics, is typically regulated by numerous genes. This study aimed to dissect the genetic regulators underlying 14 plant architectural traits and 10 biomass yield traits in the Sorghum Association Panel across two growing seasons. We identified 321 associated loci through genome-wide association studies (GWAS), involving 234 264 single nucleotide polymorphisms (SNPs). These loci include genes with known associations to biomass traits, such as maturity, dwarfing (Dw), and leafbladeless1, as well as several uncharacterized loci not previously linked to these traits. We also identified 22 pleiotropic loci associated with variation in multiple phenotypes. Three of these loci, located on chromosomes 3 (S03_15463061), 6 (S06_42790178; Dw2), and 9 (S09_57005346; Dw1), exerted significant and consistent effects on multiple traits across both growing seasons. Additionally, we identified three genomic hotspots on chromosomes 6, 7, and 9, each containing multiple SNPs associated with variation in plant architecture and biomass yield traits. Chromosome-wise correlation analyses revealed multiple blocks of positively associated SNPs located near or within the same genomic regions. Finally, genome-wide correlation-based network analysis showed that loci associated with flowering, plant height, leaf traits, plant density, and tiller number per plant were highly interconnected with other genetic loci influencing plant architectural and biomass yield traits. The pyramiding of favorable alleles related to these traits holds promise for enhancing the future development of bioenergy sorghum crops.

Singh, Anuradha (ORCID:0000000197149095)

Integrated Multiport Conductive and Wireless Architecture for Electric Vehicle Charging

This paper explores the opportunity of integrating conductive and wireless electric vehicle charging architectures and presents an integrated multiport charging architecture suitable for simultaneous conductive and wireless charging of electric vehicles. The architecture successfully integrates the primary side power electronic high frequency inverter and the high frequency transformer with the wireless transmitter and receiver to achieve independent power flow at each port. Design of the integrated magnetics is discussed, and results are presented that demonstrate the decoupled power flow at each port. The proposed integrated architecture increases the utilization of the same charging infrastructure without a proportionate increase in the cost.

Mukherjee, Subho [ORNL] (ORCID:0009000672297925)

Mapping Spiking Neural Networks to Heterogeneous Crossbar Architectures using Integer Linear Programming

Advances in novel hardware devices and architectures allow Spiking Neural Network (SNN) evaluation using ultra-low power, mixed-signal, memristor crossbar arrays. As individual network sizes quickly scale beyond the dimensional capabilities of single crossbars, networks must be mapped onto multiple crossbars. Crossbar sizes within modern Memristor Crossbar Architectures (MCAs) are determined predominately not by device technology but by network topology; more, smaller crossbars consume less area thanks to the high structural sparsity found in larger, brain-inspired SNNs. Motivated by continuing increases in SNN sparsity due to improvements in training methods, we propose utilizing heterogeneous crossbar sizes to further reduce area consumption. This approach was previously unachievable as prior compiler studies only explored solutions targeting homogeneous MCAs. Our work improves on the state-of-the-art by providing Integer Linear Programming (ILP) formulations supporting arbitrarily heterogeneous architectures. By modeling axonal interactions between neurons, our methods produce better mappings while removing inhibitive a priori knowledge requirements. We first show a 16.7-27.6% reduction in area consumption for square-crossbar homogeneous architectures. Then, we demonstrate 66.9-72.7% further reduction when using a reasonable configuration of heterogeneous crossbar dimensions. Next, we present a new optimization formulation capable of minimizing the number of inter-crossbar routes. When applied to solutions already near-optimal in area, an 11.9-26.4% routing reduction is observed without impacting area consumption. Finally, we present a profile-guided optimization capable of minimizing the number of runtime spikes between crossbars. Compared to the best-area-then-route optimized solutions, we observe a further 0.5-14.8% inter-crossbar spike reduction while requiring 1–3 orders of magnitude less solver time.

Pohl, Devin [ORNL] (ORCID:0009000040149027)

Surrogate Neural Architecture Codesign Package (SNAC-Pack)

Neural Architecture Search is a powerful approach for automating model design, but existing methods struggle to accurately optimize for real hardware performance, often relying on proxy metrics such as bit operations. We present Surrogate Neural Architecture Codesign Package (SNAC-Pack), an integrated framework that automates the discovery and optimization of neural networks focusing on FPGA deployment. SNAC-Pack combines Neural Architecture Codesign's multi-stage search capabilities with the Resource Utilization and Latency Estimator, enabling multi-objective optimization across accuracy, FPGA resource utilization, and latency without requiring time-intensive synthesis for each candidate model. We demonstrate SNAC-Pack on a high energy physics jet classification task, achieving 63.84% accuracy with resource estimation. When synthesized on a Xilinx Virtex UltraScale+ VU13P FPGA, the SNAC-Pack model matches baseline accuracy while maintaining comparable resource utilization to models optimized using traditional BOPs metrics. This work demonstrates the potential of hardware-aware neural architecture search for resource-constrained deployments and provides an open-source framework for automating the design of efficient FPGA-accelerated models.

Weitz, Jason [UC, San Diego] (ORCID:00090004631535

Spent nuclear fuel receipt rate analysis within an integrated waste management system (IWMS) architecture that includes consolidated storage

A key parameter in analyzing the performance of an integrated waste management system (IWMS) architecture for the disposition of spent nuclear fuel (SNF) is the SNF receipt rate from reactor and other custodian sites. Receipt rate in this paper means how much SNF is accepted per year for transport in the IWMS from such sites. Introducing one or more federal consolidated interim storage facilities (CISFs) into the IWMS architecture can potentially accelerate the receipt rate profile over time relative to system architectures without a CISF. This raises the question of what an optimal SNF receipt rate profile for an IWMS architecture might be in view of practical constraints and desired system performance attributes and associated metrics. This paper describes a sensitivity study on SNF receipt rates and the associated results for a selected set of IWMS scenarios aimed at informing near-term planning for interim storage capabilities and transportation assets. Two different strategies for CISF operation while awaiting availability of a disposal system to receive SNF are compared: one that relatively quickly fills an initial CISF and then idles the transportation system; and another that aims for more continuous use of transportation assets and receipt capabilities at the CISF. This study examines cost considerations and other factors, such as the timing of clearing reactor sites of SNF, efficient use of capital assets, and some other metrics that might be important to a CISF host community. Based on the analysis, an initial approach is presented that targets a continuous receipt strategy while maintaining the flexibility to step up receipt capabilities to a reasonable degree when needed and beneficial, within overall system constraints.

11 - NUCLEAR FUEL CYCLE AND FUEL MATERIALS

Spent Nuclear Fuel Receipt Rate Analysis within an Integrated Waste Manage-ment System Architecture that Includes Consolidated Interim Storage

A key parameter in analyzing the performance of an integrated waste management system (IWMS) architecture for the disposition of spent nuclear fuel (SNF) is the SNF receipt rate from reactor and other custodian sites. The introduction of one or more federal consolidated interim storage facilities (CISFs) into the IWMS architecture can enable the receipt rate profile as a function of time to be accelerated relative to system architectures without a CISF. The question then arises as to what an optimal SNF receipt rate profile for an IWMS architecture might be in view of practical constraints and desired system performance attributes and associated metrics. This paper describes a sensitivity study on SNF receipt rates and the associated results for a selected set of IWMS scenarios aimed an in-forming near-term planning for interim storage capabilities and transportation assets. Two different strategies are compared, one that fills an initial CISF quickly and then idles the transportation system while a disposal system is prepared, and a second strategy that aims to provide a more continuous use of transportation assets and receipt capabilities at the IWMS while the disposal system is readied for SNF receipt. Cost considerations and other factors such as impact on timing of clearing reactor sites of SNF, efficient use of capital assets, and other metrics, including those which may be important to a CISF host community, are examined. Based on the analysis, an initial approach is presented targeting a continuous receipt strategy while having the flexibility to step up receipt capabilities to a reasonable degree when needed and beneficial within overall system constraints.

11 - NUCLEAR FUEL CYCLE AND FUEL MATERIALS

Surrogate Neural Architecture Codesign Package (SNAC-Pack)

Neural architecture search (NAS) is a powerful approach for automating model design, but existing methods often optimize for accuracy alone or rely on proxy metrics such as bit operations (BOPs) that correlate poorly with hardware cost. This gap is particularly large for FPGA deployment, where cost is dominated by a multi-dimensional budget of lookup tables, DSPs, flip-flops, BRAM, and latency. We present the Surrogate Neural Architecture Codesign Package (SNAC-Pack), an open-source AutoML framework for hardware-aware neural architecture codesign and end-to-end FPGA deployment. SNAC-Pack runs a multi-objective global search with Optuna and NSGA-II, loading trials to a shared SQLite store that enables parallel workers across compute nodes. A hardware surrogate model outputs per-trial resource and latency estimates, avoiding the synthesis cost that would otherwise dominate the search loop. A local search stage then applies quantization-aware training (QAT) together with iterative magnitude pruning in a combined compression loop, after which the final model is synthesized to FPGA firmware via the hls4ml Python library. A YAML configuration and an optional agentic frontend let users run the pipeline on new datasets without modifying the framework. We demonstrate SNAC-Pack on jet classification at the Large Hadron Collider and superconducting qubit readout, discovering compact architectures that match or exceed strong baselines on the task metric while reducing FPGA resource utilization and, in the qubit readout case, reducing the design space exploration process from months of manual fine-tuning to hours of automated search.

Weitz, Jason [UC, San Diego]

An adaptive and stability-promoting layerwise training approach for sparse deep neural network architecture

This work presents a two-stage adaptive framework for progressively developing deep neural network (DNN) architectures that generalize well for a given training data set. In the first stage, a layerwise training approach is adopted where a new layer is added each time and trained independently by freezing parameters in the previous layers. We impose desirable structures on the DNN by employing manifold regularization, sparsity regularization, and physics-informed terms. We introduce a ε – δ – stability-promoting concept as a desirable property for a learning algorithm and show that employing manifold regularization yields a ε – δ stability-promoting algorithm. Further, we also derive the necessary conditions for the trainability of a newly added layer and investigate the training saturation problem. In the second stage of the algorithm (post-processing), a sequence of shallow networks is employed to extract information from the residual produced in the first stage, thereby improving the prediction accuracy. Numerical investigations on prototype regression and classification problems demonstrate that the proposed approach can outperform fully connected DNNs of the same size. Moreover, by equipping the physics-informed neural network (PINN) with the proposed adaptive architecture strategy to solve partial differential equations, we numerically show that adaptive PINNs not only are superior to standard PINNs but also produce interpretable hidden layers with provable stability. As a result, we also apply our architecture design strategy to solve inverse problems governed by elliptic partial differential equations.

42 ENGINEERING

Conserved Macromolecular Architecture of Poplar Secondary Cell Walls Revealed by ssNMR and Atomistic Modeling

The macromolecular architecture of plant secondary cell walls governs wood's mechanical and biochemical properties, yet its natural intra-species variability remains poorly characterized. Here, we combined 13C solid-state NMR (ssNMR), multivariate statistical analysis, and molecular modeling to profile nanoscale structure across 13 genetically diverse Populus trichocarpa genotypes grown in 13C-enriched atmospheres. SsNMR-derived phenotypes spanning composition, structure, mobility, and inter-polymer proximities reveal a conserved architecture, with a subtle yet coordinated variation organizing into dominant structural and secondary mobility axes. A representative atomistic model captures these features and reproduces experimental metrics. Molecular dynamics simulations support a weak but consistent positive correlation between cellulose abundance and crystalline-like order, with interior cellulose chains enriched in tg (trans-gauche) conformations without expanding crystalline cores. Together, experiment and simulation reveal a genetically buffered, broadly conserved nanoscale architecture across genotypes, where subtle fine-tuning of cellulose bundling and matrix packing balances mechanical performance with biological function.

09 BIOMASS FUELS

Neural architecture codesign for fast physics applications

We develop a pipeline to streamline neural architecture codesign for physics applications to reduce the need for ML expertise when designing models for novel tasks. Our method employs neural architecture search and network compression in a two-stage approach to discover hardware efficient models. This approach consists of a global search stage that explores a wide range of architectures while considering hardware constraints, followed by a local search stage that fine-tunes and compresses the most promising candidates. We exceed performance on various tasks and show further speedup through model compression techniques such as quantization-aware-training and neural network pruning. We synthesize the optimal models to high level synthesis code for FPGA deployment with the hls4ml library. Additionally, our hierarchical search space provides greater flexibility in optimization, which can easily extend to other tasks and domains. We demonstrate this with two case studies: Bragg peak finding in materials science and jet classification in high energy physics, achieving models with improved accuracy, smaller latencies, or reduced resource utilization relative to the baseline models.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS

The architecture of resilience: a genome assembly of Myrothamnus flabellifolia sheds light on desiccation tolerance and sex determination

Myrothamnus flabellifolia is a dioecious resurrection plant endemic to southern Africa that has become an important model for understanding desiccation tolerance. Despite its ecological and medicinal significance, genomic and transcriptomic resources for the species are limited. We generated a chromosome-level, haplotype-resolved reference genome assembly and annotation for M. flabellifolia and conducted transcriptomic profiling across a natural dehydration–rehydration time course in the field. Genome architecture and sex determination were characterized, and co-expression network and cis-regulatory element (CRE) enrichment analyses were used to investigate dynamic responses to desiccation. The 1.28-Gb genome exhibits unusually consistent chromatin architecture with unique chromosome organization across highly divergent haplotypes. We identified an XY sexual system with a small sex-determining region on Chromosome 8. Transcriptomic responses varied with dehydration severity, pointing to early suppression of growth, progressive activation of protective mechanisms, and subsequent return to homeostasis upon rehydration. Late embryogenesis abundant and early light-induced protein transcripts were dynamically regulated and showed enrichment of abscisic acid and stress-responsive CREs pointing toward conserved responses. Together, this study provides foundational resources for understanding the genomic architecture and reproductive biology of M. flabellifolia and offers new insights into the mechanisms of desiccation tolerance.

chromosome structure

Architecture and performance of Perlmutter's 35 PB ClusterStor E1000 all-flash file system

NERSC's newest system, Perlmutter, features a 35 PB all-flash Lustre file system built on HPE Cray ClusterStor E1000. Here, we present its architecture, early performance figures, and performance considerations unique to this architecture. We demonstrate the performance of E1000 OSSes through low-level Lustre tests that achieve over 90% of the theoretical bandwidth of the SSDs at the OST and LNet levels. We also show end-to-end performance for both traditional dimensions of I/O performance (peak bulk-synchronous bandwidth) and nonoptimal workloads endemic to production computing (small, incoherent I/Os at random offsets) and compare them to NERSC's previous system, Cori, to illustrate that Perlmutter achieves the performance of a burst buffer and the resilience of a scratch file system. Finally, we discuss performance considerations unique to all-flash Lustre and present ways in which users and HPC facilities can adjust their I/O patterns and operations to make optimal use of such architectures.

97 MATHEMATICS AND COMPUTING

Architectural Approaches for Integrating ADMS and DERMS: Challenges, Comparisons, and Real-World Use Cases

The electrical distribution landscape is rapidly transforming due to the proliferation of distributed energy resources (DERs) such as solar panels, wind turbines, battery storage systems, combined heat and power units, and electric vehicles, introducing variability and uncontrollability that traditional grid operators are ill-equipped to manage. This transformation is further accelerated by advancements in Information and Communication Technology infrastructure that connects control centers with end devices, demanding automation and a deeper understanding of new technologies by utility personnel. Advanced grid control techniques using system-level optimization, Artificial Intelligence, and Machine Learning at the enterprise level and distributed level are evolving to address these issues. There is also an opportunity to utilize the enormous data created by these new DER technologies in the grid. Advanced Distribution Management Systems (ADMS) and Distributed Energy Resource Management Systems (DERMS) are critical in addressing these challenges by automating grid operations and enhancing reliability. Given the relatively recent development of ADMS and DERMS, and the still relatively low level of ADMS and DERMS deployment in the industry, there is a notable deficiency in the comprehensive understanding of the challenges and benefits associated with these new technologies, especially with their complementary natures and integration architectures. This paper aims to bridge the knowledge gap in ADMS and DERMS integration, presenting three distinct integration architectures currently available, and discussing the challenges and benefits of each architecture to guide utilities, industry professionals, and researchers in optimizing grid management and decision-making processes for a resilient and efficient energy future.

24 POWER TRANSMISSION AND DISTRIBUTION

Mesoporous Thin Film Architectures: Addressing Material Demands through Molecular Self-Assembly

Mesoporous thin films spark interest across a wide range of disciplines due to their tunable nanostructures, large internal surface areas, and strong compatibility with planar optical, electronic, and microfluidic devices. While attention in the porous materials community has shifted toward macroporous or disordered nanoporous systems, a resurgence in mesoporous thin film research is underway, driven by new molecular self-assembly methods, advanced materials chemistry, and improved characterization techniques. The integration of high-χN block copolymer design, kinetically persistent micelle templating, and postdeposition processing protocols now allows control over structural parameters such as pore size, wall thickness, porosity, and connectivity. These advances have overcome many of the thermodynamic and processing constraints that previously limited widespread adoption. Rather than serving only as high-surface-area supports, mesoporous thin films are engineered as active interfaces where responsive chemistries and nanoscale confinement act in tandem. Embedding switchable ligands, thermoresponsive polymers, redox mediators, or ion-selective groups directly within the pore walls enables real-time control over transport, optical, and electrochemical properties. These capabilities open up new directions in adaptive coatings, gated membranes, and fast-response biosensors. To further expand their functional scope, mesoporous films are integrated into hierarchical and multicomponent architectures. Techniques such as triblock terpolymer templating, crack-directed assembly, and nanoimprint lithography allow for control over spatial organization on the micron and submicron scale and pore system orientation. This enables programmable anisotropy, enhanced molecular diffusion, and wavelength-selective photonic behavior, essential for next-generation sensing, catalysis, and energy applications. Such structural and functional complexity requires equally sophisticated characterization. Multimodal and in situ techniques can track material dynamics under operational conditions. Recent progress includes extended-range ellipsometric porosimetry (EP) for hierarchical architectures, vacuum EP for interface energetics, time-resolved EP for diffusion kinetics, and correlative AFM-SAXS mapping. The introduction of advanced neutron-based spectroscopies, particularly quasielastic neutron scattering (QENS), promises to provide real-time access to ion transport dynamics and segmental motion under nanoscale confinement, offering a path toward deeper mechanistic understanding of structure-performance correlations in mesoporous systems. This Account reflects the technical advances made and the interdisciplinary collaborations that have shaped our collective vision. The particular dimensions of mesopores enable us to subtly tune interactions at the molecular, interfacial, and mesoscopic levels that permit us to harness nanoconfinement. What emerges is a versatile, modular platform capable of chemical gating, energy transduction, and sensing with a level of tunability unmatched by other porous materials. We highlight critical challenges including the need for more robust large-area processing, a deeper understanding of dynamic behavior under cycling, and better integration with device-level architectures. Our strategies support the transition of mesoporous thin films into active high-performance components in next-generation energy, environmental, and biomedical systems.

oxides

Understanding the Influence of Chain Architecture on the Transport Quantities of Polymer Electrolytes with Covalently Bonded Anions

Here, we use a combination of experiments and coarse-grained molecular dynamics simulations to elucidate the structure–property relationships in polymer electrolytes obtained by the copolymerization of poly(vinyl ethylene carbonate─lithium styrene bis(trifluoromethanesulfonyl)imide) or p(VEC-LiSTFSI). Experiments show that the conductivity reduces with increasing anion (i.e., STFSI) fraction on the chain, and the cation transference number (t + ) is found to be dependent on the anion fraction. Furthermore, a significant fraction of unpolymerized VEC monomers are observed. Since it is inherently difficult to experimentally control the chain architecture and the amount of unpolymerized VEC in these systems, we perform coarse-grained molecular dynamics simulations on model polymer systems with different chain architectures to mimic the plausible experimental systems. Specifically, we look at the differences in transference numbers arising from (i) a random copolymer of VEC and STFSI monomers; (ii) a blend of VEC-STFSI copolymer with VEC monomers; and (iii) a ternary blend of the VEC homopolymer, STFSI homopolymers, and VEC monomers. The ternary blend model demonstrates the closest resemblance with the experimental transference numbers and diffusivities. The lithium diffusivity obtained from the coarse-grained models with VEC monomers (plasticizers) is about 1.5 times that of the model without VEC monomers, showing that the plasticizing effect of VEC monomers is modest. We rationalize the experimental observations based on aggregate and cluster analyses obtained from molecular simulations. This work reveals that polymer electrolyte chain architecture and plasticizers can critically influence the transport properties, and these parameters should be considered when designing single ion conducting polymeric electrolytes.

cluster distribution

Investigating the Role of Polymer Architecture in Poly(methyl methacrylate) Depolymerization

Advanced architectures in polymers have garnered traction within the last two decades due to their distinct and tunable properties compared to linear analogs. However, the effects of architecture on nascent polymethacrylate depolymerization strategies remain underexplored. Herein, we investigate the depolymerization behavior of poly(methacrylate)-based star copolymers synthesized via a core-first approach. By incorporating either labile chain-end or pendent-group triggers, we demonstrate the first direct comparison of bulk depolymerization behaviors in star versus linear methacrylate copolymers under matched conditions. While pendent-triggered systems required higher loadings of N-(methacryloxy)phthalimide methacrylate (PhthMA) to achieve similar mass loss compared to linear copolymers, chain-end-initiated depolymerization showed enhanced efficiency of monomer liberation in the star topology. Furthermore, these findings highlight the importance of considering macromolecular architecture in designing sustainable polymers and provide actionable guidelines for installing depolymerization triggers based on polymer topology.

Copolymers