Integrated Simulation of PIP-II at Fermilab
We describe progress towards a community software ecosystem for efficient modeling of the Fermilab PIP-II complex for design validation and virtual test stands.
SEARCH · Engineering Papers
Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.
Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.
We describe progress towards a community software ecosystem for efficient modeling of the Fermilab PIP-II complex for design validation and virtual test stands.
Material Testing 2.0 (MT2.0) is a paradigm that advocates for the use of rich, full-field data, such as from digital image correlation and infrared thermography, for material identification. By employing heterogeneous, multi-axial data in conjunction with sophisticated inverse calibration techniques such as finite element model updating and the virtual fields method, MT2.0 aims to reduce the number of specimens needed for material identification and to increase confidence in the calibration results. To support continued development, improvement, and validation of such inverse methods—specifically for rate-dependent, temperature-dependent, and anisotropic metal plasticity models—we provide here a thorough experimental data set for 304L stainless steel sheet metal. The data set includes full-field displacement, strain, and temperature data for seven unique specimen geometries tested at different strain rates and in different material orientations. Commensurate extensometer strain data from tensile dog bones is provided as well for comparison. We believe this complete data set will be a valuable contribution to the experimental and computational mechanics communities, supporting continued advances in material identification methods.
The National Fire Protection Association, with support from the Department of Energy, executed a multi-year initiative to develop, enhance, and disseminate Distributed Energy Resources Safety Training (DERST) tools for U.S. emergency responders. As Distributed Energy Resources (DER)—such as solar photovoltaics, battery energy storage systems (ESS), electric vehicles (EVs), and associated infrastructure—become increasingly prevalent, the NFPA identified a critical need for up-to-date standardized, accessible, and effective safety training tailored for the fire service and related public safety professionals. The project delivered a comprehensive suite of educational resources to improve responders’ abilities to safely manage DER-related incidents. This included: • Revised Modular Training Courses: Updated classroom-based DER safety courses, now modular and accessible nationwide through fire academies and the North American Fire Training Directors (NAFTD) network. • Live Burn Testing & Research: A full-scale controlled burn of a DER-equipped residential structure provided real-world data and insights, forming the basis for updated best practices. • A Gamified Simulation Tool – Firefighters Incident Response Simulation Tool (FIRST): A first-of-its-kind, multiplayer, scenario-based simulation using the Unreal Engine 5.0 to train responders in a realistic virtual, multi-DER incident environment. • Field Familiarization Software Tools & Prop Guide: Digital DER field familiarization evolutions software guide and a prop development manual to support field-based DER training exercises, enhancing responders' hands-on familiarity with DER infrastructure and collaboration on virtual incident responses. • National Dissemination Strategy: Strategic partnerships with NAFTD, Vector Solutions, and others enabled wide-scale distribution, with over 5,000 departments accessing resources and 1,100+ departments adopting the simulator in the first seven months. Also provided a web portal for easy access to all training and simulation programs developed under this grant for the U.S. responder community. Key findings from the project—particularly from the burn test—led to paradigm shifts in fire response tactics. For example, traditional approaches to garage fires may be hazardous if DERs are present, due to explosive off gassing and thermal runaway risks. The new training emphasizes scene assessment, stand-off approaches, thermal imaging verification, and careful post-incident cooling of DER components to prevent reignition. This initiative has had a significant national impact, raising awareness, enhancing preparedness, and supporting safer DER incident response practices. Significant engagement from the media, public safety organizations, and PBS coverage has further amplified the reach and adoption of NFPA’s DER safety training, tools, and simulations.
Conventional and recently developed approaches for estimating turbulent scalar fluxes under stable atmospheric conditions are evaluated, with a focus on gases for which fast sensors are not readily available. First, the relaxed eddy accumulation (REA) classical approach and a recently proposed mixing length parameterization, labeled A22, are tested against eddy-covariance computations. Using high-frequency measurements collected from two contrasting sites (the frozen tundra near Utqiaġvik, Alaska, and a sparsely vegetated grassland in Wendell, Idaho, during winter), it is shown that the REA and A22 models outperform the conventional Monin–Obukhov similarity theory (MOST) utilized widely to infer fluxes from mean gradients. Second, scenarios where slow trace gas sensors are the only viable option in field measurements are investigated using digital filtering applied to fast-response sensors to simulate their slow-response counterparts. With a filtered scalar signal, the observed filtered eddy-covariance fluxes are referred to here as large-eddy-covariance (LEC) fluxes. A virtual eddy accumulation (VEA) approach, akin to the REA model but not requiring a mechanical apparatus to separate the gas flows, is also formulated and tested. A22 outperforms VEA and LEC in predicting the observed unfiltered (total) eddy-covariance (EC) fluxes; however, VEA can still capture the LEC fluxes well. This finding motivates the introduction of a sensor response time correction into the VEA formulation to offset the effect of sensor filtering on the underestimated net averaged fluxes. The only needed parameter for this correction is the mean velocity at the instrument height, a surrogate of the advective timescale. The VEA approach is very suitable and simple to use with gas sensors of intermediate speed (∼ 0.5 to 1 Hz) and with conventional open- or closed-path setups.
Harnessing virtual power plants enhances the integration of distributed energy resources into utility grids for a sustainable energy future. Virtual power plants (VPPs) aggregate DERs to enhance resource adequacy and reduce emissions. U.S. utilities are exploring various technologies to manage DERs effectively. FERC Order 2222 allows DERs to participate in both wholesale and retail markets. Enhancing observability and controllability of behind-the-meter (BTM) DERs is essential for reliable grid operations. A hierarchical control architecture can improve coordination among residential energy resources. Field tests showed nearly 20% energy savings and 30% peak power reduction during grid events. Effective management of DERs requires enhanced situational awareness to prevent grid congestion. Integrating DER management systems (DERMS) with existing planning tools can improve operational security. Near-real-time grid models can validate optimal resource set points against resource uncertainty. Traditional uninterruptible power supplies (UPS) can be upgraded to support grid services and become part of VPPs. Upgrading UPS systems can reduce costs by 75% and unlock significant battery capacity. New battery management systems and grid-aware controllers are essential for optimizing UPS performance. Continued research and development are necessary to address challenges in integrating DERs into utility grids. Encouraging customer participation in pilot programs is vital for the evolution of VPPs. Here, the shift towards price-responsive DERs and VPPs is expected to enhance energy distribution efficiency.
Accurate identification of material parameters is crucial for predictive modeling in computational mechanics. Here, the two primary approaches in the experimental mechanics community for calibration from full-field digital image correlation data are known as finite element model updating (FEMU) and the virtual fields method (VFM). In VFM, the objective function is a squared mismatch between internal and external virtual work or power. In FEMU, the objective function quantifies the weighted mismatch between model predictions and corresponding experimentally measured quantities of interest. It is minimized by iteratively updating the parameters of an FE model. While FEMU is seen as more flexible, VFM is commonly used instead of FEMU due to its considerably greater computational expense. However, comparisons between the two methods usually involve approximations of gradients or sensitivities with finite difference schemes, thereby making direct assessments difficult. Hence, in this study, we compare VFM and FEMU in the context of numerically-exact sensitivities obtained through local sensitivity analyses and the application of automatic differentiation software. To this end, we conduct a series of test cases to assess both methods under practical challenges using a finite strain elastoplasticity model.
An overview of the Solenoidal Large Intensity Device (SoLID) and its scientific program will be given in this talk. SoLID is a spectrometer/detector system proposed to exploit the full potential of the Jefferson Lab (JLab) 12 GeV energy upgrade. SoLID will push the limit of luminosity frontier in hadronic physics with its unique capability to handle very high rates with large acceptance under high luminosity (1037-39/cm2/s). A rich and vibrant scientific program has been developed for SoLID, including but not limited to the precision study of the 3d nucleon structure in both momentum space using Semi-Inclusive Deep Inelastic Scattering (SIDIS) and coordinate space using Deep Virtual Exclusive Reactions (DVER), probing physics beyond the Standard Model with Parity Violating Deep Inelastic Scattering (PVDIS), and investigating the gluonic field contribution to the proton structure and proton mass via J/¿ threshold production. The SoLID collaboration has developed a robust, low risk and flexible conceptual design, with a base line design capable of accomplishing its scientific goals and flexibility to adopt the cutting-edge technology. Detector subsystems have been tested with prototypes in realistic high luminosity conditions and are demonstrated to function well under extremely challenging environment to satisfy the requirements of planned experiments.
Steel Thread is a NA-22 venture that seeks to build trustworthy, reliable AI models that can be used in a wide variety of nonproliferation tasks. A key aspect of building these models is developing appropriate benchmarks and evaluation methods, which will enable the venture to identify and adapt models to provide the most value in the nonproliferation domain. Benchmarks must be relevant to key tasks in this domain, such as question answering, information retrieval, document summarization and classification, consensus analysis, and image and data analysis. This report 1) provides an overview of benchmark design, evaluation, and challenges; 2) reviews a variety of open benchmarks, with a focus on language models and tasks; and 3) identifies benchmarks that are most relevant to Steel Thread. This report is intended to serve as a basis for further efforts to classify and evaluate benchmarks and their correlation with success on nonproliferation-specific tasks. The Steel Thread venture has defined benchmarks to be a particular combination of a dataset (or datasets) and a metric (or metrics) conceptualized as representing one or more specific tasks or sets of abilities for a specific modality. It is adopted by a research community as a shared framework for comparing methods.1 It includes 1) Data: Labeled (a designated subset not used for training, which could be all the data), 2) Metric: A way to quantify performance, 3) Task/Ability: What the benchmark is testing, 4) Protocol: A structured and repeatable evaluation process, 5) Baseline/Reference Model: For comparison; could be statistical, rule-based, SME-derived, or another model, and 6) Maintenance Plan: to update with new information over time; important for long-term utility. For further clarity, the definition includes what a benchmark, in this context, is not. It is not a corpus of training data, specific to a model (it is intended to apply to a range of models), a universal evaluation of performance, a guarantee that the ‘top’ model on the leaderboard will be the best fit for every specific use case, an all-encompassing proof of a model’s universal quality, nor is it a one-size-fits-all measure of success. It does not cover every real-world constraint (like operational, ethical, or cost considerations), a systems integration test, or a unit test. This definition was inspired by and resulted from discussions within the Steel Thread Benchmarking Task Force. This group was formed to define what we would mean as a benchmark within Steel Thread but persisted as the need to develop a thorough understanding of the large and expanding existing benchmarking space. This technical report is a result of the group’s divide and conquer approach to exploring this space. The release of benchmarks might not be progressing as quickly as model development, but it is moving very fast, as many benchmarks quickly become saturated, when state-of-the-art models score so close to the benchmark’s ceiling that their results are virtually indistinguishable. At that point, the test no longer differentiates between new systems, so researchers usually stop reporting scores as the benchmark no longer informs about improvements from the next generation of models. In the OpenAI announcement of GPT-5, they reported results on six flagship public benchmarks (AIME 2025, SWE-bench Verified, Aider Polyglot, MMMU, HealthBench Hard, GPQA) but the full system-card covers roughly thirty-five separate evaluations, comprising hundreds of test task items in total. There have been some efforts to summarize benchmarks in specific fields, like for text-to-image generation, but these surveys have had a narrow methodology scope. Therefore, a comprehensive survey of all benchmarks or even all benchmarks that could be relevant to Steel Thread is outside of the scope of this report. We chose some specific benchmarks to investigate in detail.
To manufacture light-weight, advanced metal alloy components for gas turbine engines, quench heat-treatment processes are typically used. By quenching the component from elevated temperatures, the alloy sometimes undergoes a solid-state phase transformation which produces special microstructures with the required, enhanced mechanical properties. However, the quenching can also lead to cracks forming in the component. Addressing the quench cracking problems adds a significant burden to the cost, schedule, and energy demand of manufacture. Currently, optimizing the quench process to mitigate or avoid the cracking is performed largely by trial-and-error, relying heavily on costly experimental (thermocouple)trials to understand the local thermal gradients which cause the cracks to form. In this first part (Phase 1) of the work, high-performance computing is employed to establish the ability of modern CFD (computational fluid dynamics) to alleviate or wholly replace the experimental quenching trials by virtual testing. A Baseline CFD model is defined and its accuracy established to be comparable to(and which usually exceeds) the accuracy of existing HTC (heat-transfer coefficient) based simulation methods of quenching. As a first-principles based approach, “calibration” of the Baseline CFD model is independent of the quench process itself, but instead relies on the accuracy of the underlying (modeled),generic two-phase fluid processes which cannot be currently resolved by CFD for large, industrial-scale cases. A novel, high-fidelity DNS capability has been developed and verified to examine and further improve upon the mean-field closure submodels on which the Baseline CFD approach is based.1
To manufacture light-weight, advanced metal alloy components for gas turbine engines, quench heat-treatment processes are typically used. By quenching the component from elevated tempera-tures, the alloy sometimes undergoes a solid-state phase transformation which produces special microstructures with the required, enhanced mechanical properties. However, the quenching can also lead to cracks forming in the component. Addressing the quench cracking problems adds a significant burden to the cost, schedule, and energy demand of manufacture. Currently, optimizing the quench process to mitigate or avoid the cracking is performed largely by trial-and-error, relying heavily on costly experimental (thermocouple) trials to understand the local thermal gradients which cause the cracks to form. In this first part (Phase 1) of the work, high-performance computing is employed to establish the ability of modern CFD (computational fluid dynamics) to alleviate or wholly replace the experimental quenching trials by virtual testing. A baseline CFD model is defined and its accuracy established to be comparable to (and which usually exceeds) the accuracy of existing HTC (heat-transfer correlation) based simulation methods of quenching. As a first-principles based approach, “calibration” of the Baseline CFD model is independent of the quench process itself, but instead relies on the accuracy of the underlying (modeled), generic two-phase fluid processes which cannot be currently resolved by CFD for large, industrial-scale cases. A novel, high-fidelity DNS capability has been developed and verified to examine and further improve upon the mean-field closure submodels on which the Baseline CFD approach is based.
SpinQuest is a cutting-edge, high-luminosity Drell-Yan experiment utilizing polarized hy- drogen and deuterium targets to measure the Sivers asymmetry for the light sea quarks in the nucleon. Detecting a nonzero Sivers asymmetry would provide clear evidence for nonzero or- bital angular momentum of sea quarks. The Sivers asymmetry presents itself as an azimuthal asymmetry in the production of virtual photons via the Drell-Yan process, and SpinQuest will be able to measure this asymmetry using the existing SeaQuest dimuon spectrometer. In addi- tion to making measurements sensitive to the sea quark Sivers function, we will also measure the azimuthal asymmetry in the production of J/ψ particles, which is sensitive to the gluon Sivers function. Additionally, observing a sign change in the Sivers asymmetry between this measurement and future measurements at the Electron-Ion Collider would be a test of a funda- mental prediction of Quantum Chromodynamics. In this poster we will review the physics and technology underpinning the experiment.
Ion energy distributions generated by pulsed laser interactions with materials are essential for applications ranging from materials science to oncology. Ion energy characterization is particularly important for an emerging mass spectrometry technique called virtual-slit cycloidal mass spectrometry (VS-CMS). The ion energy distribution influences the design and performance of VS-CMS instruments, as well as the efficacy of laser-driven ionization methods in various fields. Several established techniques, including the retarding potential method, time-of-flight (TOF) analysis, and electrostatic energy analyzers, have been employed to measure ion energy distributions. The wide range of ion energies reported highlights the strong dependence of ion energy on laser parameters, target materials, and experimental conditions, as well as the necessity of making independent measurements of the ion energy distribution for specific laser systems and materials. This paper presents the design and characterization of a simple TOF-based apparatus for measuring ion energy distributions from pulsed laser ionization without external fields. This approach minimizes perturbation of electron-ion dynamics and enables simultaneous energy measurements at multiple spatial positions. Here, the apparatus was tested using a nanosecond pulsed Nd:YAG laser operating at 1064 nm, 532 nm, and 266 nm on solid copper sheets at various laser fluences. Simultaneous measurements at different distances provide new insights into ion-electron interactions post-ionization and demonstrate the influence of laser wavelength and fluence on ion energy distributions.
Building automation and controls are becoming increasingly complex with the emergence of Grid Integrated Efficient Buildings (GEBs) as well as new highly efficient sequences of operation and data-driven control schemes. However, there remains a significant gap in hands-on training opportunities for building operators and technicians to gain practical experience with advanced control systems in a low-risk environment. This paper presents BOPTEST (Building Optimization Performance Test) as a suitable platform for workforce training in building controls and GEB technologies. BOPTEST provides a suite of standardized building simulation test cases with a REST API, real-time control interfaces through BACnet, semantic models connecting users to building data, and built-in calculation of control metrics and performance indicators. The platform enables trainees to interact with virtual buildings using industry-standard protocols while learning how to implement and innovate control strategies. The training platform is designed to offer a structured and interactive learning experience for building engineers, helping them effectively develop, learn, and retain skills in fault identification, troubleshooting, and correction. The workflow is divided into three main phases: 1) Setup, 2) Exercise, and 3) Review, each comprising specific activities performed by either the instructor or the student. Initial pilot training sessions have yielded positive feedback from instructors and participants and demonstrates that BOPTEST effectively fills an industry need for a low-risk training resource via simulation of real building control systems, allowing trainees to gain practical experience before working in the field. The platform's ability to provide immediate performance feedback while maintaining familiar industry interfaces makes it particularly suitable for workforce development programs. This work provides a replicable model for leveraging building simulation in control education and training.
Building automation and controls are becoming increasingly complex with the emergence of Grid Integrated Efficient Buildings (GEBs) as well as new highly efficient sequences of operation and data-driven control schemes. However, there remains a significant gap in hands-on training opportunities for building operators and technicians to gain practical experience with advanced control systems in a low-risk environment. This paper presents BOPTEST (Building Optimization Performance Test) as a suitable platform for workforce training in building controls and GEB technologies. BOPTEST provides a suite of standardized building simulation test cases with a REST API, real-time control interfaces through BACnet, semantic models connecting users to building data, and built-in calculation of control metrics and performance indicators. The platform enables trainees to interact with virtual buildings using industry-standard protocols while learning how to implement and innovate control strategies. The training platform is designed to offer a structured and interactive learning experience for building engineers, helping them effectively develop, learn, and retain skills in fault identification, troubleshooting, and correction. The workflow is divided into three main phases: 1) Setup, 2) Exercise, and 3) Review, each comprising specific activities performed by either the instructor or the student. Initial pilot training sessions have yielded positive feedback from instructors and participants and demonstrates that BOPTEST effectively fills an industry need for a low-risk training resource via simulation of real building control systems, allowing trainees to gain practical experience before working in the field. The platform's ability to provide immediate performance feedback while maintaining familiar industry interfaces makes it particularly suitable for workforce development programs. This work provides a replicable model for leveraging building simulation in control education and training.
This project addresses several key barriers to implement the next generation demand response applications and provides a clear understanding of implementing hierarchical and standalone control using AMI data. Through this program, Eaton has developed and tested a meter-as-a-controller prototype with the help of other partners--- National Renewable Energy Laboratory (NREL), Electric Power Research Institute (EPRI), Pecan St Inc. (PSI), and Delaware Electric Cooperative (DEC). The controller can utilize residential controllable loads such as heating, ventilation, and air conditioner (HVAC), electric water heater and distributed energy resources like solar PV and battery energy storage systems for off-setting the demand that is required from the grid, thus providing reliable grid-services for demand reduction or peak shaving. The controller is also capable of coordinating the resources of the premises for better management and energy efficiency while meeting the comfort bound of the premises owner as quality-of-service. The development has been demonstrated in a three virtual-home setup at system performance lab of NREL with real appliances (HVAC, electric water heater, solar PV, and battery). The technology has also been proved through laboratory and field demonstration with successful interconnectivity (e.g., end-to-end communication and data exchange) between the residential appliances and utility through the RF network at Delaware Electric Co-op (DEC) in Delaware.
Explorations into the internal dynamics of hadrons are constantly evolving, and the requirement for experimental results to verify theoretical models of hadron structure is paramount. A key area in this field is the study of Generalised Parton Distributions (GPDs), which are functions used to model the momenta of quarks and gluons within hadrons, and the methods to access GPDs experimentally. One such scattering process that allows access to these is Timelike Compton Scattering. TCS complements existing Deeply Virtual Compton Scattering experiments and allows investigation into the universality of GPDs through access to the real and imaginary parts of the parton helicity independent GPD Hq via beam spin asymmetries (BSA), and it provides novel access to the real and imaginary parts of the parton helicity dependent GPD ˜Hq through target polarisation asymmetries (TSA). This thesis work presents a comparative study with the first published BSA for TCS at the Thomas Jefferson National Accelerator Facility (JLab), alongside a first time extraction of a Target Spin Asymmetry with the Summer 2022 data taking run. JLab hosts the Continuous Electron Beam Accelerator Facility (CEBAF) which provides a 12 GeV electron beam to four experimental halls. Hall-B contains the CEBAF Large Acceptance Spectrometer, which took data across three run periods on a longitudinally polarised NH3 and ND3 fixed target from 2022-2023, to extract measurements of electron-proton scattering, from which a TCS signal could be extracted. The thesis discusses work done to understand and eliminate contributions from the non-/low-polarised nuclear background, testing pre-established cuts to eliminate pion background from a dilepton (e+e-) final state and modifying them as needed for the new experimental run, and attempts to hone in on a clean TCS signal from which to extract the two asymmetry observables. A comparison with existing BSA results was performed; however, the statistical errors are too large to draw a significant conclusion as to whether there is agreement across each bin. More data is needed for a multidimensionally binned extraction. A proof of principle was achieved in the TSA measurements, with two out of four kinematic bins showing preliminary agreement in shape with theoretical values. Again, the errors are significant due to the contributions from the nuclear background. To support these conclusions, a further study was done, which takes into account an estimate of the asymmetries with the full available dataset (this thesis is based only on data taken in the summer set; at the time of writing processing was still being conducted for the final two datasets), as well as an estimate including additional future experiment days that were awarded in July 2024. Additional work was done on a secondary project exploring the feasibility of measuring TCS at the upcoming Electron Ion Collider, supporting the design proposal for the detector for the first interaction region and giving a positive outlook for the future of these types of measurements beyond JLab.
We describe AthenaK: a new implementation of the Athena++ block-based adaptive mesh refinement framework using the Kokkos programming model. Finite volume methods for Newtonian, special relativistic, and general relativistic (GR) hydrodynamics and magnetohydrodynamics (MHD), and GR-radiation hydrodynamics and MHD, as well as a module for evolving Lagrangian tracer or charged test particles (e.g., cosmic rays) are implemented using the framework. In two companion papers, we describe (1) a new solver for the Einstein equations based on the Z4c formalism, and (2) a GRMHD solver in dynamical spacetimes also implemented using the framework, enabling new applications in numerical relativity. By adopting Kokkos, the code can be run on virtually any hardware, including CPUs, GPUs from multiple vendors, and emerging Advanced RISC Machine processors. AthenaK shows excellent performance and weak scaling, achieving over 1 billion cell updates per second for hydrodynamics in three dimensions on a single NVIDIA Grace Hopper processor. It does this with a typical parallel efficiency of 80% on 65,536 AMD GPUs on the OLCF Frontier system. Such performance portability enables AthenaK to leverage modern exascale computing systems for challenging applications in astrophysical fluid dynamics, numerical relativity, and multimessenger astrophysics.
The NEAMS Multiphysics Applications team continues to assess code usability and functionality for microreactor design and safety analyses, while demonstrating that NEAMS tools capture both steady-state and transient behavior across distinct microreactor concepts. In FY2025, the team advanced full-core, high-fidelity, multiphysics models that solve more complex problems and strengthen verification/validation for several microreactor systems: heat-pipe microreactor (HPMR), gas-cooled microreactor (GCMR), and the KRUSTY experiment. These models employ the MOOSE MultiApp/Transfers architecture with Griffin for neutronics, BISON for heat conduction/thermomechanics, Sockeye for heat pipes, SAM/THM for coolant channels and loops, and SWIFT for hydride behavior, with meshes generated via the MOOSE Reactor Module. The graphite models available in the Grizzly code were also investigated for future analyses. For the HPMR, a Na-HPMR variant was constructed to align with recently validated heat-pipe experiments and Sockeye’s LCVF capability, enabling mechanistic heat-pipe transients and startup modeling. The Na-HPMR will serve as the primary model for HPMR investigations in upcoming tasks. The load-following and single heat-pipe failure scenarios (Griffin/BISON/Sockeye), which were previously modeled for the K-HPMR, were replicated for the Na-HPMR, showing strong negative temperature feedback and highly localized thermal effects, respectively, while the startup case captured vapor-front progression and heat-removal activation. Solid mechanics was added to the previously built K-HPMR full-core model in BISON, showing minimal impact on steady-state reactivity yet enabling stress-field predictions that prepare the path for full-core TRISO performance analyses. For the GCMR, automated steady-state and four transient scenarios were executed using Griffin/BISON/SAM/SWIFT. Results confirm robust inherent safety: power collapses promptly in loss-of-cooling events, the inlet-temperature drop settles to a new equilibrium, and a single-channel blockage yields only a ~30 K local fuel-temperature rise with <0.4% power decrease. SWIFT-predicted hydrogen redistribution affects reactivity during both steady-state and transient conditions, underscoring its importance. A Brayton-cycle balance of plant (BOP) model in SAM/THM demonstrated stable startup behavior, and xenon-driven reactivity during load following was analyzed. To improve TRISO-compact temperature fidelity, a fast multiscale Heat Source Decomposition (HSD) treatment was implemented. Against heterogeneous benchmarks, HSD reduces underprediction of kernel temperatures and lowers predicted peak powers in reactivity-insertion transients compared to previous homogenized models. KRUSTY warm-critical validation progressed from FY2024 baselines: the 15Ȼ insertion shows excellent agreement in peak power (~2% high) and temperature trends, and the 30Ȼ case was automated via a feedback controller that maintained power near 3 kW for ~150 s with close agreement to data. The successful modeling of the warm critical tests has laid a strong foundation for simulating more complex nuclear system tests in the years ahead. Throughout FY2025, developer feedback was provided (e.g., MOOSE batch mesh generation, distributed pre-split meshes, Griffin sweeper on displaced meshes), several new models were contributed to the Virtual Test Bed, and an OECD-NEA WPRS multiphysics benchmark based on the HPMR was initiated to enable broader cross-comparison and best-practice development with the nuclear community at large.