Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “concurrent solutions”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

ECP libraries and tools: An overview

The Exascale Computing Project (ECP) Software Technology and Co-Design teams addressed the growing complexities in high-performance computing (HPC) by developing scalable software libraries and tools that leverage exascale system capabilities. As we enter the exascale era, the need for reusable, optimized software solutions that can handle the unique challenges posed by these systems becomes increasingly important. The primary challenges the ECP teams faced were to create software libraries and tools that are performant on exascale architectures and portable and usable across diverse hardware platforms. Efforts addressed issues related to concurrent execution, memory management, and the integration of heterogeneous computing resources, such as GPUs from multiple vendors. The ECP’s strategy involved a structured development process encompassing the creation, optimization, and deployment of software in collaboration with industry, academia, and national laboratories. The project was organized into several technical areas: co-design of domain-specific suites with target applications, programming models and runtimes, development tools, mathematical libraries, data and visualization tools, and software ecosystem and delivery mechanisms. ECP has successfully developed a large portfolio of software libraries and tools that demonstrate significant improvements in performance and scalability on exascale systems. These products have been integrated into the Department of Energy’s computing facilities, supporting various scientific applications and ensuring robust performance across different hardware setups. ECP advancements in software development for exascale computing highlight the importance of a collaborative and adaptive approach to handling next-generation HPC systems complexities. The lessons learned emphasize the need for continuous engagement with end-users and vendors, and the importance of maintaining a balance between innovation and practical implementation. Future efforts will focus on ensuring scalability, keeping pace with rapid hardware advancements, and further enhancing the interoperability and usability of the software ecosystem. In conclusion, subsequent articles in this special issue provide in-depth discussions and case studies into specific library and tool efforts.

97 MATHEMATICS AND COMPUTING↗

Special Considerations for the Removal and Disposal of Micro-Reactor Experiments

Idaho National Laboratory (INL) is preparing to host several microreactor experiments through 2030 and beyond. Two new test beds currently under development will operate as microreactor or nuclear system experiment user facilities. Test bed experiments will be performed in series, with each nuclear experiment installed, operated, and removed before installation of the subsequent experiment. The necessarily brief transition period between experiments introduces unique equipment removal and radioactive waste disposal challenges. This paper evaluates equipment removal and radioactive waste management topics associated with tight sequencing of nuclear experiments and presents some of the solutions and approaches currently planned to meet these special considerations for one such experiment, the Molten Chloride Reactor Experiment (MCRE), slated for operation at INL’s Laboratory for Operation and Testing in the United States (LOTUS). The MCRE project is a collaboration between Southern Company Services, TerraPower, and INL, among others, to provide integral nuclear data that will advance molten salt fast reactor technology. MCRE will be the first critical fast-spectrum circulating fuel system ever operated and the first experiment operated in the LOTUS test bed. The experiment will nominally operate at zero power with planned low power excursions as part of operational planning for MCRE has been to minimize at-power operations to limit fission product formation while still achieving experimental objectives. After the experiment is complete, MCRE will be allowed to radioactively decay for a short period (nominally 90 days) prior to system defueling, flushing, removal, and disposal of all MCRE equipment, readying the test bed for the next nuclear experiment. Equipment removal and radioactive waste disposal have been integral to MCRE project planning since the Cooperative Research and Development Agreement was formalized in 2021. Due to requirements for future use of the test bed, the MCRE system, of necessity, must be removed in a much shorter time frame than typical for historic reactor decommissioning projects at INL. This results in minimal time for radioactive decay, resulting in not only elevated radiation dose rates but also the presence of short-to-medium-lived isotopes not typically encountered in the reactor decommissioning and radioactive waste management space. Additional unique constraints placed on the project include lack of intrinsic remote-operations capabilities in the test bed, space constraints in the test bed once MCRE has been installed, and contamination minimization requirements to return the test bed to as-found conditions to enable future use. This paper discusses planned solutions to these challenges. Approaches for implementing remote or semi-remote technologies in a non-hot cell environment with limited space availability are discussed. The paper also summarizes the systems engineering approach for concept development and design of equipment removal systems, which are being implemented concurrent with the MCRE design phase, providing feedback to system designers to incorporate features enabling efficient and safe equipment removal approaches. The timing of this paper at a relatively early phase of the project is meant to highlight the importance of early planning for nuclear system decommissioning while reactor design is ongoing to allow design feedback on componentry driven from decommissioning system needs. As new, innovative nuclear reactor technologies enter the nuclear market sector, reactor experiments are crucial for providing new integral nuclear data that establish safe operational margins for technology advancement. Microreactor experiments at INL will require safe, effective, and timely decommissioning approaches, which in turn require nuclear systems designed to expedite decommissioning. In addition to the advancement of nuclear reactor technologies, these nuclear experiments provide an opportunity to demonstrate, deploy, and test new removal and disposal capabilities to support the next generation of nuclear reactor technology.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

GeantV: Results from the Prototype of Concurrent Vector Particle Transport Simulation in HEP

Full detector simulation was among the largest CPU consumers in all CERN experiment software stacks for the first two runs of the Large Hadron Collider. In the early 2010s, it was projected that simulation demands would scale linearly with increasing luminosity, with only partial compensation from increasing computing resources. The extension of fast simulation approaches to cover more use cases that represent a larger fraction of the simulation budget is only part of the solution, because of intrinsic precision limitations. The remainder corresponds to speeding up the simulation software by several factors, which is not achievable by just applying simple optimizations to the current code base. In this context, the GeantV R&D project was launched, aiming to redesign the legacy particle transport code in order to benefit from features of fine-grained parallelism, including vectorization and increased locality of both instruction and data. This paper provides an extensive presentation of the results and achievements of this R&D project, as well as the conclusions and lessons learned from the beta version prototype.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

2-Aminoimidazole reduces fouling and improves membrane performance

Biofouling is difficult to control and hinders the performance of membranes in all applications but is of particular concern when natural waters are purified. Fouling, via multiple mechanisms (organic-only, biofouling-only, cell-deposition-only, and organic + biofouling), of a commercially available membrane (control) and a corresponding membrane coated with an anti-biofouling 2-aminoimidazole (2-AI membrane) was monitored and characterized during the purification of a natural water. Results show that the amount of bacterial cell deposition and organic fouling was not significantly different between control and 2-AI membranes; however, biofilm formation, concurrent or not with other fouling mechanisms, was significantly inhibited (95–98%, p < 0.001) by the 2-AI membrane. The limited biofilm that formed on the 2-AI membrane was weaker (as indicated by the polysaccharide to protein ratio) and thus presumably easier to remove. The conductivity rejection by the 2-AI and control membranes was not significantly different throughout the 75-h experiments, but the rejection of dissolved organic carbon by biofouled (biofouling-only, cell-deposition-only, and organic + biofouling) 2-AI membranes was statistically higher (10–12%, p = 0.003–0.07). When biofouled, the water permeance of the 2-AI membranes decreased significantly less (p < 0.05) over 75 hours than that of the control membranes, whether or not other additional types of fouling occurred concurrently. Despite the initially lower water permeances of 2-AI membranes (11% lower on average than controls), the 2-AI membranes outperformed the controls (10–11% higher average water permeance) after biofilm formation occurred. Overall, 2-AI membranes fouled less than controls without detriment to water productivity and solute rejection.

59 BASIC BIOLOGICAL SCIENCES↗

Towards robust laser beam propagation in atmospheric turbulence

High-fidelity optical propagation through the atmosphere is essential for free-space optical technologies, including laser-based remote sensing and optical communication. However, atmospheric turbulence severely distorts beams and compromises system performance. In this work, we employ hypergeometric-Gaussian (HyGG) vortex beams as probes to characterize and mitigate atmospheric turbulence. Using over 250,000 experimental and simulated frames, we show that refining the power spectrum density (PSD) can reduce numerical prediction errors by up to 79.8%. Concurrently, experimental observations supported by numerical simulations demonstrate that HyGG beams exhibit superior turbulence resilience across multiple metrics compared to conventional Gaussian beams, particularly in their ability to withstand over 5 times stronger turbulence while maintaining similar intensity fluctuations. These dual investigations, on both turbulence mitigation and robust beam solutions, converge to form a unified strategy for enhancing free-space optical system performance. Collectively, our findings provide new insights into light–turbulence interactions and highlight the practical utility of vortex beams under atmospheric conditions.

Zhang, Boyu↗

$\mathrm{RADICAL}$-Pilot and $\mathrm{PMIx}$/$\mathrm{PRRTE}$: Executing Heterogeneous Workloads at Large Scale on Partitioned $\mathrm{HPC}$ Resources

Execution of heterogeneous workflows on high-performance computing (HPC) platforms present unprecedented resource management and execution coordination challenges for runtime systems. Task heterogeneity increases the complexity of resource and execution management, limiting the scalability and efficiency of workflow execution. Re-source partitioning and distribution of tasks execution over portioned re-sources promises to address those problems but we lack an experimental evaluation of its performance at scale. Here this paper provides a performance evaluation of the Process Management Interface for Exascale (PMIx) and its reference implementation PRRTE on the leadership-class HPC plat-form Summit, when integrated into a pilot-based runtime system called RADICAL-Pilot. We partition resources across multiple PRRTE Distributed Virtual Machine (DVM) environments, responsible for launching tasks via the PMIx interface. We experimentally measure the work-load execution performance in terms of task scheduling/launching rate and distribution of DVM task placement times, DVM startup and termination overheads on the Summit leadership-class HPC platform. Integrated solution with PMIx/PRRTE enables using an abstracted, standardized set of interfaces for orchestrating the launch process, dynamic process management and monitoring capabilities. It extends scaling capabilities allowing to overcome a limitation of other launching mechanisms (e.g., JSM/LSF). Explored different DVM setup configurations provide insights on DVM performance and a layout to leverage it. Our experimental results show that heterogeneous workload of 65,500 tasks on 2048 nodes, and partitioned across 32 DVMs, runs steady with resource utilization not lower than 52%. While having less concurrently executed tasks resource utilization is able to reach up to 85%, based on results of heterogeneous workload of 8200 tasks on 256 nodes and 2 DVMs.

97 MATHEMATICS AND COMPUTING↗

Classical Benchmarks for Variational Quantum Eigensolver Simulations of the Hubbard Model

Simulating the Hubbard model is of great interest to a wide range of applications within condensed matter physics, however its solution on classical computers remains challenging in dimensions larger than one. The relative simplicity of this model, embodied by the sparseness of the Hamiltonian matrix, allows for its efficient implementation on quantum computers, and for its approximate solution using variational algorithms such as the variational quantum eigensolver. While these algorithms have been shown to reproduce the qualitative features of the Hubbard model, their quantitative accuracy in terms of producing true ground state energies and other properties, and the dependence of this accuracy on the system size and interaction strength, the choice of variational ansatz, and the degree of spatial inhomogeneity in the model, remains unknown. Here we present a rigorous classical benchmarking study, demonstrating the potential impact of these factors on the accuracy of the variational solution of the Hubbard model on quantum hardware, for systems with up to 32 qubits. We find that even when using the most accurate wavefunction ansätze for the Hubbard model, the error in its ground state energy and wavefunction plateaus for larger lattices, while stronger electronic correlations magnify this issue. Concurrently, spatially inhomogeneous parameters and the presence of off-site Coulomb interactions only have a small effect on the accuracy of the computed ground state energies. Our study highlights the capabilities and limitations of current approaches for solving the Hubbard model on quantum hardware, and we discuss potential future avenues of research.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Does Water Enhance Mg Intercalation in Oxides? The Case of a Tunnel Framework

The presence of H 2 O has been linked to enhancements in the reactivity of cathodes for Mg 2+ electrochemistry. If the enhancements were mimicked by nonaqueous solvents, they could enable Mg batteries with transformational energy density. However, the extent to which H 2 O may boost actual intercalation of Mg 2+ , as opposed to competing reactions, has not been elucidated. Here, in this paper, we evaluate its role as additive in the electrochemistry of a tunnel polymorph of V 2 O 5 in a nonaqueous Mg 2+ electrolyte. The electrochemical response and V reduction in the cathodes positively correlated with H 2 O concentration, but it was not concurrent with commensurate changes in cell volume and Mg content. These observations indicate that H 2 O does not enhance Mg 2+ intercalation, but rather, it promotes competing pathways. This work shows the importance of accurately probing reactions in multivalent electrolytes. Importantly, it indicates that H 2 O is not a universal solution to the challenge of Mg 2+ intercalation in oxides.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Coupled feldspar dissolution and secondary mineral precipitation in batch systems: 6. Labradorite dissolution, calcite growth, and clay precipitation at 60 °C and pH 8.2–8.4

Here, we conducted experiments on concurrent labradorite dissolution, calcite precipitation, and clay precipitation in batch reactor systems and tracked reaction processes using multiple isotope tracers. Labradorite was chosen for its role as a major and reactive component in basalt; the experiments thus directly impact our understanding of CO 2 storage in basalt aquifers and enhanced rock weathering. We doped initial solutions with 29 Si, 43 Ca, and Ca 13 CO 3 (s). Experiments were conducted at 60 °C and pH ~ 8.3 for up to 840 h, with isotope ratios in the experimental aqueous solutions measured using MC-ICP-MS. Unidirectional rates of labradorite dissolution near equilibrium were approximately two orders of magnitude slower than far-from-equilibrium rates reported in the literature. Calcite growth occurred near equilibrium and the rates were limited by the labradorite dissolution rates. In the steady state phase, the interplay of these three heterogeneous reactions—labradorite dissolution, calcite growth, and clay precipitation—results in a coupled system that approaches a near-equilibrium state. The system does not reach true equilibrium because labradorite continues to dissolve, albeit at a much slower rate near equilibrium. The overall reaction can be approximated as, Na 0.4 Ca 0.6 Al 1.6 Si 2.4 O 8 + 0.6HCO 3 - + 1·.7H 2 O + 0.4H + → 0.4Na + + 0.6CaCO 3(s) + 0.5Al 2 Si 2 O 5 (OH) 4(s) + 0.6Al(OH) 4 - + 1.4SiO 2 o (aq). The experimental results show that using short-term far-from-equilibrium rate constants would lead to an overestimation of feldspar weathering rates at the Earth’s surface (e.g., basalt weathering and enhanced rock weathering) and CO 2 mineralization in basalt aquifers.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Microchannel-based Membrane-less Extraction of Li from Unconventional Lithium Sources & the Separation of REE

This final report provides an overview of the Project's entire duration, covering July 1, 2021 to December 31, 2023. It primarily focuses on the achievements, technological developments, and unique challenges the team faced while working on separating and extracting Lithium from produced waters. The project's primary aim was to create an integrated, high-throughput, membrane-less, and modular microfluidic platform that could extract Lithium from unconventional sources. We have successfully met all goals and milestones envisioned in the SOPO document. The most critical primary milestones, including the Go-No-Go milestone (refer to the Gantt chart in the Appendices), were successfully accomplished. We demonstrated phase separation (>90%) and extraction (>85%) performance in the MPSE using synthetic, and representative produced water composition feed at 50 ml/min total flow through MPSE 36. We have also performed a parametric study of the MPSE operations, beyond the scope of SOPO, exploring operating conditions of current and broader interest. The extended investigation of operational parameters is concurrent with our efforts to seek further development of the MPSE technology beyond the scope of the Project. Along these lines of development, we have made efforts to be responsive to DOE calls for technological developments of other types of resources (beyond PW) for the recovery of Critical Materials and higher TRL development (beyond TRL 4). During the work on this Project, we developed and implemented three innovative technical approaches that emerged from our efforts to successfully meet the Project milestones. The innovative & original technical approaches developed and implemented in this Project are now the contributions to process engineering that could be clearly credited to the Project. First, Convergent Design Approach is a comprehensive feedforward & feedback loop of four design phases: i) design for functionality, ii) design for manufacturing, iii) design for sustainability, and iv) design for market. Next was Process Intensification. A major aim of this Project was to create an innovative phase separation & extraction microscale-based technology for Li separation – thus the words microchannel-based in the Project title. A microscale-based technology is intrinsically in the center of the Process Intensification domain as defined by its unique principles. Therefore, Process Intensification was implicitly envisioned in the Project’s SOPO. Lastly, Time Scale Analysis is a novel tool for discovering the needs and directions of Process Intensification implementations in any process technology. This Project is fully credited for developing and implementing the three novel technical approaches mentioned above. These are general contributions to process engineering that emerged from this Project. Beyond the original SOPO scope, the OSU-U.Pitt research group utilized a Convergent Design methodology, integrating first-principles mathematical modeling with experimental validation on the Minimum Development Vehicle. By creating these Digital Twins, the team rapidly assessed manufacturing iterations to support TEA analysis. This framework further enabled the development of advanced Surface Modification Techniques, where hydrophobic and oleophobic coating strategies were optimized via Digital Twin tools and validated through rigorous 100-hour longevity testing. TEA Analysis: The closing efforts of this Project were focused on the TEA analysis. TEA analysis had two primary functions: i) enabling critical assessments of design variations withing 10 the Concurrent Design Approach, thus enabling evolution of the MPSE design to reach faster- better-cheaper alternatives; and ii) to create a bridge between the accomplishments of this Project and future projects of higher TRL, beyond TRL 6 level. It is important to note that the TEA model created in the Project stirred the technological solutions for the recovery of critical materials toward a vision of a very profitable modular plant that has unique zero-waste water discharge signature. More importantly, thanks to our experimental performance data and conservative assumptions, the TEA model predicts minimal technological and investment risks. Low cost of a modular unit of a nominal capacity of [1000 tons of Li 2 CO 3 /year] positions the MPSE based technology within the reach of community investors, thus offering a paradigm shift in the development of critical technologies. The project successfully navigated two primary challenges: solvent selection and manufacturing adaptation. Restricted by the SOPO to existing literature for lithium recovery, the team identified a critical need for a "material excellence program" to develop next-generation solvents, eventually concluding with a preliminary investigation into promising Ionic Liquids (ILs). Simultaneously, COVID-19 supply chain disruptions forced a pivot from traditional manufacturing to advanced additive methods at ATAMI-OSU. By transitioning from stainless steel to 3D-printed polymer substrates, the team achieved a transformative three-order-of- magnitude reduction in manufacturing costs and compressed prototyping timelines from several months to just two days. The MPSE technology offers significant energy, environmental, and economic advantages by overcoming the traditional bottlenecks of phase-separation hardware and contactor size. Unlike conventional mixer-settlers or membrane-based systems, MPSE operates without moving parts or fouling-prone membranes, achieving robust performance even with challenging, viscous, or particulate-heavy feeds. Key performance metrics include an energy intensity reduction of 5–50x (3–40 kJ/m 3 ) compared to incumbent technologies and a dramatic reduction of processing time to under 60 seconds, which drastically reduces the physical plant footprint. These technical efficiencies translate into superior economic outcomes; for a 100 t/year Li 2 CO 3 facility, implementing MPSE is projected to nearly halve contactor CAPEX (from $\$$6.08M to $\$$3.01M) and significantly increase the project's Net Present Value (NPV), derisking new investment and enabling distributed critical-mineral processing configurations. The commercialization of MPSE technology is being spearheaded by Vigsur Dynamics Inc., which has adopted a structured, parallel approach to technical and business development since its formation in January 2026. Following extensive customer discovery and engagement with the Oregon State University accelerator, Vigsur Dynamics is working to establish a business model that transitions from pilot demonstrations to modular hardware sales, ultimately aiming for a "build-own-operate" service strategy. Current technical milestones—including 100 hours of continuous operation, superior energy efficiency, and successful 6-unit modular scale-up— provide a foundation for this transition. Backed by ongoing IP licensing and a growing network of industrial and venture advisors, the company is actively de-risking the platform to replace conventional mixer-settler systems in the critical minerals market.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

New Mechanism for Yield Point Phenomena

Yield point phenomena (YPP) are widely attributed to discrete dislocation locking by solute atmospheres. An alternate YPP mechanism was recently suggested by simulations of Ta single crystals without any influence of solutes or discrete dislocations. The general meso-scale (GM) simulations consist of crystal plasticity (CP) plus accounting for internal stresses of geometrically necessary dislocation content. GM predicted the YPP while CP did not, suggesting a novel internal stress mechanism. The predicted YPP varied with crystal orientation and boundary conditions, contrary to expectations for a solute mechanism. The internal stress mechanism was probed by experimentally deforming oligocrystal Ta samples and comparing the results with independent GM simulations. Strain distributions of the experiments were observed with high-resolution digital image correlation. A YPP stress–strain response occurred in the 0–2% strain range in agreement with GM predictions. Shear bands appeared concurrent with the YPP stress–strain perturbation in agreement with GM predictions. At higher strains, the shear bands grew at progressively slower rates in agreement with GM predictions. It was concluded that the internal stress mechanism can account for the existence of YPP in a wide variety of materials including ones where interstitial-dislocation interactions and dislocation transient avalanches are improbable. Furthermore, the internal stress mechanism is a CP analog of various micro-scale mechanisms of discrete dislocations such as pile-up or bow-out. It may operate concurrently with strain aging, or either mechanism may operate alone. A suggestion was made for a future experiment to answer this question.

42 ENGINEERING↗

Climate Challenges and Nonproliferation: Addressing the Issue through Technical Cooperation

Climate change is an urgent global challenge that is already severely impacting many regions of the world. The 2015 Paris Agreement was a landmark achievement, while the 2023 UN COP 28 Climate Conference in Dubai recognized nuclear energy and its applications as a proven and sustainable means to help societies adapt to climate change and take measures to mitigate and, possibly, reverse its effects. Twenty-two countries, supported by over 120 companies, pledged to triple the share of global nuclear energy generation by 2050. The global interest in nuclear applications for peaceful uses has never been higher. The deployment of advanced reactors around the globe – focused primarily on lowering the carbon footprint and allowing for socio-economic development – must be addressed responsibly through setting priorities for its implementation. Tackling the challenges that new technologies, such as advanced reactors, bring to the nonproliferation regime is at the top of the list. The regime stands on the three pillars of the Nuclear Nonproliferation Treaty (NPT): nonproliferation, peaceful uses of nuclear energy, and disarmament. Recognizing the inherent risks of expanding applications of nuclear materials and technologies for peaceful uses, all such efforts must concurrently strengthen the nonproliferation norm enshrined in the NPT. The challenges of adapting to and mitigating climate change with the use of advanced reactors is already impacting the discussions and expectations about the future of the NPT. The dialog with non-traditional domestic, and international partners on a strategic technical cooperation approach is key to proposing and implementing scientific and technical solutions for urgent climate issues, while framing the discussion within nonproliferation requirements and commitments, and demonstrating the underlying value of the NPT in support of peaceful nuclear applications. This paper addresses adaptation to climate challenges and mitigation of them through advanced reactors and establishes nonproliferation linkages that derive from the deployment of this technology. This paper presents an assessment of the benefits of mechanisms for technical cooperation and peaceful uses of nuclear technology in the framework of Article IV of the NPT.

Prah, Christina↗

A phenology- and trend-based approach for accurate mapping of sea-level driven coastal forest retreat

The rapid replacement of upland forest by encroaching marshland is a striking manifestation of global sea-level rise (SLR). Timely and high-resolution information on the location and extent of transition forest (the ecotone between upland forest and marsh where tree mortality due to seawater intrusion begins) is fundamental to understanding the processes and patterns of SLR-driven landscape reorganization. Despite its significance, accurate characterization of salt-impacted transition forest remains challenging due to the complexity of coastal environments, scarcity of ground-truth data, and the lack of effective mapping algorithms. Here we use the full archive of Landsat images between 1984 and 2021 to investigate the spectral, temporal, and phenological characteristics of transition forest, and develop a robust framework for monitoring coastal vegetation shifts in the mid-Atlantic U.S., a global SLR hotspot. Here, we found that transition forest exhibits strong negative NDVI trends and a deviation of land surface phenology from marsh and upland forest that distinguishes itself from surrounding vegetation. By integrating temporal trends and land surface phenology, our results demonstrate superior discrimination between marsh and coastal forests to existing map products (e.g. NOAA Coastal Change Analysis Program, National Land Cover Database) that allows a reliable identification of the coastal treeline. We applied the approach to map regional land cover in 1985, 2000 and 2020 (overall classification accuracy >92%) and found that the area of coastal forest decreased by 22.0% from 1985 to 2020, the majority of which transitioned to marshland (92.3%, 5.3 × 10 3 ha). Based upon fine-scale patterns of coastal transgression, we created a practical workflow for spatially explicit quantification of forest retreat rates. Concurrent with rising sea level, coastal forests migrated upslope from 0.63 (± 0.27) m above sea level in 1985 to 0.78 (± 0.32) m above sea level in 2020, and horizontal forest retreat rates accelerated from 3.1 (range of 0–36) m yr -1 during 1985–2000 to 4.7 (0–55) m yr -1 during 2001–2020. As SLR continues to accelerate, our study may serve as a scalable solution for consistent tracking of coastal landscape evolution that is urgently needed for sustainable forest and wetland management.

54 ENVIRONMENTAL SCIENCES↗

Sensitivity analysis of an automated fault detection algorithm for residential air-conditioning systems

The state of the art of fault detection and diagnosis (FDD) for residential air-conditioning systems is expensive and not yet amenable to widespread implementation. FDD for homes can significantly reduce utility costs, and increase the lifespan of the equipment. The cost barriers currently, however, make FDD for homes economically unviable for large scale implementation. In prior work, we offered a solution to reduce FDD costs by proposing an automated fault detection algorithm to serve as a screening step before more expensive FDD tests can be conducted. The algorithm uses only the home thermostat and local weather information to identify thermodynamic parameters and detect high-impact air-conditioning faults, including those that occur during equipment installation. We had tested the algorithm on a single EnergyPlus™ model of a home in Orlando, Florida. The thermodynamic parameter identification process is highly nonconvex involving several local optimal solutions. In this paper we propose a novel method to select the best model for fault detection from among the list of local optimal solutions to make the algorithm more robust to homes of different construction, without which the fault detection process would be infeasible. Another unique contribution of the paper is implementing the solution on real-world data. We also bring the algorithm closer to market by testing it on real-world data. We implement the algorithm on data obtained from experiments conducted by the Florida Solar Energy Center (FSEC) on a laboratory home equipped with a heat pump where faults were intentionally added for a period of seven months. The algorithm successfully detected an undercharge fault with 70.6% accuracy, concurrent duct leakage and undercharge faults with 85.2% accuracy, and duct leakage faults with 69.1% accuracy. A sensitivity analysis is also performed on EnergyPlus models of nine types of homes that vary in construction to demonstrate the robustness of the algorithm. Finally, the algorithm achieves an average accuracy of 71% for no-fault condition, 77% for 40% undercharge fault, and 76% for duct-leak fault.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

FitCache: A Transparent Drop-In Framework for Multi-Tier Caching to Accelerate Distributed Deep Learning Workloads

Training in Deep learning (DL) remains highly compute- and data-intensive, with I/O becoming a critical bottleneck as models and datasets scale. Recent studies report that data loading can dominate training time, especially on large-scale HPC systems with shared parallel file systems (PFS). Existing caching approaches either rely on single-tier designs or require intrusive modifications to training pipelines, limiting their portability and effectiveness. In this work, we present FitCache, a transparent drop-in framework for multi-tier caching to accelerate distributed DL training by coordinating fast local memory (e.g., DRAM, Persistent Memory (PMem)) and NVMe as hierarchical caches atop PFS. Our design adapts to hardware diversity, i.e., if NVMe is missing, memory transparently acts as a caching tier, ensuring stable performance. FitCache transparently intercepts I/O requests and issues concurrent fetches across all tiers, returning data from the fastest responder without centralized metadata or static redirection paths. FitCache adapts to dynamic workloads and heterogeneous clusters while maintaining POSIX compatibility. Experiments on Frontier (2048 GPUs) and smaller research clusters show that FitCache reduces training time by up to 40% and per-batch I/O latency by up to 71.6% compared to Lustre Orion PFS, offering a drop-in solution for scalable DL training.

Hu, Guangxing [ORNL] (ORCID:0009000283203614)↗

A Comprehensive Review and Qualitative Analysis of Micro-Combined Heat and Power Modeling Approaches

Concurrent production of electrical and thermal energy from a Combined Heat and Power (CHP) device is an attractive tool to address the growing energy needs of the planet. Micro CHP (µCHP) systems can reduce a building’s primary energy consumption, reduce carbon footprint, and enhance resiliency. Modeling of the µCHP helps understand the system from multiple perspectives and helps discover errors earlier, improves impact analysis and simulation of system solutions for ease of integration with the building. Consequently, there is a need for analysis of the impact of µCHP modeling approach on its reliability and flexibility. The primary objective of this paper is to review the state-of-the art models in the µCHP space with a focus towards internal combustion engine as the primary mover (PM) and limit the study to system modeling, calibration, and validation methodologies. Based on the analysis, recommendations for further model considerations and refinements are presented.

42 ENGINEERING↗

Fingerprinting shock-induced deformations via diffraction

Abstract During the various stages of shock loading, many transient modes of deformation can activate and deactivate to affect the final state of a material. In order to fundamentally understand and optimize a shock response, researchers seek the ability to probe these modes in real-time and measure the microstructural evolutions with nanoscale resolution. Neither post-mortem analysis on recovered samples nor continuum-based methods during shock testing meet both requirements. High-speed diffraction offers a solution, but the interpretation of diffractograms suffers numerous debates and uncertainties. By atomistically simulating the shock, X-ray diffraction, and electron diffraction of three representative BCC and FCC metallic systems, we systematically isolated the characteristic fingerprints of salient deformation modes, such as dislocation slip (stacking faults), deformation twinning, and phase transformation as observed in experimental diffractograms. This study demonstrates how to use simulated diffractograms to connect the contributions from concurrent deformation modes to the evolutions of both 1D line profiles and 2D patterns for diffractograms from single crystals. Harnessing these fingerprints alongside information on local pressures and plasticity contributions facilitate the interpretation of shock experiments with cutting-edge resolution in both space and time.

36 MATERIALS SCIENCE↗

High-Level Synthesis of Parallel Specifications Coupling Static and Dynamic Controllers

The increased need for efficient ways to implement domain-specific accelerators is driving design methodologies towards the use of abstractions higher than the Register Transfer Level (RTL). In this scenario, High Level Synthesis (HLS) plays a significant role by enabling the automatic generation of custom hardware accelerators starting from high level descriptions (e.g., C code). Conventional HLS tools exploit parallelism mostly at the Instruction Level (ILP). They statically schedule the input specifications, and build centralized Finite State Machine (FSM) controllers. However, aggressive exploitation of ILP in many applications has diminishing returns and, usually, centralized approaches do not efficiently exploit coarser parallelism because FSMs are inherently serial. In this paper we present a HLS framework able to synthesize applications that, beside ILP, also expose Task Level Parallelism (TLP). An application can expose TLP through annotations that identify the parallel functions (i.e., tasks). To generate accelerators that efficiently execute concur- rent tasks, we need to solve several issues: devise a mechanism to support concurrent execution flows, exploit memory parallelism, and manage synchronization. To support concurrent execution flows, we introduce a novel adaptive controller. The adaptive controller is composed of a set of interacting control elements that independently manage the execution of a single operation or function call. These control elements check dependencies and resource constraints at runtime, enabling as soon as possible execution. To support parallel access to shared memories and synchronization, we introduce a novel Hierarchical Memory Interface (HMI). With respect to previous solutions, the proposed interface supports multi-ported memories and atomic memory operations, which commonly occur in parallel programming. Our framework can generate the hardware implementation of C functions by employing two different approaches, depending on its characteristics. If a function exposes TLP, then the framework generates hardware implementations based on the adaptive controller. Otherwise, the framework implements the function by exploiting a more conventional FSM approach, which is optimized for ILP exploitation. We evaluate our framework on a set of parallel applications, and show substantial performance improvements (average speedup of 4.7) with limited area over- heads (average area increase of 5.48 times).

Castellana, Vito G.↗