Engineering PapersSearch

SEARCH · Engineering Papers

Results for “model checking”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

A Bayesian approach to time-domain photonic Doppler velocimetry analysis

Photonic Doppler velocimetry (PDV) is an established technique for measuring the velocities of fast-moving surfaces in high-energy-density experiments. In the standard approach to PDV analysis, the short-time Fourier transform (STFT) is used to generate a spectrogram from which the velocity history of the target is inferred. The user chooses the form, duration, and separation of the window function. Here, in this study, we present a Bayesian approach to infer the velocity directly from the PDV oscilloscope trace, without using the spectrogram for analysis. This is clearly a difficult inference problem due to the highly periodic nature of the data, but we find that with carefully chosen prior distributions for the model parameters, we can accurately recover the injected velocity from synthetic data. We validate this method using PDV data collected at the STAR two-stage light gas gun at Sandia National Laboratories, recovering shock-front velocities in quartz that are consistent with those inferred using the STFT-based approach and are interpolated across regions of low signal-to-noise data. Although this method does not rely on the same user choices as the STFT, we caution that it can be prone to misspecification if the chosen model is not sufficient to capture the velocity behavior. Analysis using posterior predictive checks can be used to establish whether a better model is required, although more complex models come with additional computational cost, often taking more than several hours to converge when sampling the Bayesian posterior. We, therefore, recommend it be viewed as a complementary method to that of the STFT-based approach.

Allison, James R. [First Light Fusion Ltd., Yarnto

An Autonomous MCP Bridge to Rucio: Enhancing Data Management Accessibility for High Energy Physics

The Rucio Data Management System [1] is an important tool used by High Energy Physics experiments, including those at Fermi National Accelerator Laboratory, to store and manage exabyte-scale scientific datasets. Despite its central role in coordinating data across globally distributed storage sites, Rucio's command line interface (CLI) presents a steep learning curve, and makes it difficult for scientists to navigate through. To solve this issue, a containerized Model Context Protocol (MCP) [2] server was built that connects Large Language Models directly to Rucio, allowing AI agents to handle data tasks by using simple, natural language rather than memorized terminal commands. The core engineering focus of this project was moving the server away from slow terminal commands that require text parsing and replacing them with a native Python Client API toolset and a planned REST API framework. Moving to the Python API handles data operations directly in memory, which helps clear up formatting errors, provides the AI with clean, structured JSON data and speeds up tool execution. To prove that the system actually works, a benchmarking pipeline was also built with various questions to test the AI across four different model configurations. The questions included finding data scopes, tracking down specific datasets, and checking replication rules. Through benchmarking, early runs showed that with raw terminal text, the model would get confused and stuck, whereas switching to the Python API to feed the AI clean, structured data yielded massive improvement. By creating an intelligent and autonomous bridge to a storage network, this project shows how AI can be implemented in scientific data management, which ultimately helps scientists at Fermilab spend less time sorting through data and more time focusing on their experiments and analysis.

Akella, Kashyap [William Rainey Harper Coll.]

Dynamic analysis of fully constrained Cable-Driven Parallel Robots for automated prefabricated component installation

This paper presents a dynamic analysis and validation framework to assess a fully constrained six-anchor Cable-Driven Parallel Robot (CDPR) for automated installation of prefabricated facade components. Compared with conventional eight-anchor systems, the six-anchor configuration simplifies setup and reduces cost, but it also reduces control authority, shrinks the wrench-feasible workspace, and tightens orientation limits. Consequently, it is unclear a priori whether dynamically feasible trajectories exist to move the end effector from pickup to the facade. A constrained trajectory optimization is formulated to enforce the system dynamics, cable-tension bounds, and pose/velocity limits, and the framework is evaluated in simulation at three levels: (i) an idealized reference model, (ii) a lab-scale prototype model incorporating measured anchor misalignments and identified damping, and (iii) a full-scale three-story building model with load decomposition for structural feasibility checks. Across these scenarios, the analysis shows that optimal, constraint-satisfying trajectories exist that move the end effector from pickup to installation while maintaining a near-plumb, level orientation at the final pose. Collectively, this multi-scale dynamic analysis and validation framework supports the deployment readiness of the six-anchor CDPR and provides a prototype-based sensitivity case study of how measured anchor placement deviations affect feasibility.

CDPR

Better practices for inferring ecosystem water use strategy from eddy covariance data

Eddy covariance data are critical for inferring ecosystem water use strategies. Yet, such inferences are sensitive to a range of assumptions applied across studies, hindering our understanding of water use strategies within and across eddy covariance sites. A recent analysis across 151 FLUXNET2015 and AmeriFlux-FLUXNET datasets found that poor model performance was the key driver of non-robust inferences of ecosystem water use strategies. Here, we leverage this previous analysis to (i) identify the specific assumptions that improve inference model performance across most sites, (ii) explain the mechanisms behind the performance improvements, and (iii) check whether better performance improves water use inference. We find that the common practice of fitting a model to canopy conductance (G c ) derived from the evapotranspiration (ET) observations, rather than to observed ET itself, artificially amplifies data errors and degrades the model performance. Next, accounting for vegetation dynamics by applying a growing season filter or incorporating satellite LAI data improves performance, but the former practice may remove soil water stress periods. Lastly, using the leaf-to-air vapor pressure deficit (VPD l ) derived from ET observations as a model input may artificially inflate performance. Based on these results, we recommend selecting observed ET (rather than derived G c ) as the response variable, carefully accounting for vegetation dynamics, and avoiding derived VPD l as a model input; these best practices improve model performance by c. 20% and robustness by c. 80% across all eddy covariance sites. Nevertheless, the performance improvements do not always correspond to more robust inference of water use strategies, as model parameter selection and surface energy budget closure corrections still strongly influence the ecosystem water use parameter estimation in a site-specific manner.

AmeriFlux

Scaling open-weight large language models for hydropower regulatory information extraction: A systematic analysis

Information extraction from regulatory and technical documents using large language models (LLMs) involves practical trade-offs between extraction quality and computational cost. We evaluate eight open-weight LLMs spanning 0.6B–70B parameters on hydropower licensing documents and report deployment-oriented evidence under a unified extraction schema and evaluation protocol. Across the model set, we observe clear scale-dependent trends in both baseline extraction quality and the effectiveness of reflective reasoning (self-checking) under our fixed-prompt, no-augmentation setting. Mid-scale models often provide a favorable balance of accuracy and efficiency, whereas the smallest models show limited or inconsistent gains from the reasoning variants tested. Larger models achieve the highest overall F1 scores but incur substantially greater compute and infrastructure requirements. We further find that reliability failure modes can distort conventional metrics in this domain: in particular, high recall can coincide with systematic extraction errors when models fabricate values for fields that are absent from the source text, underscoring the importance of conservative null handling and evidence-grounded evaluation. Overall, our study provides a reproducible resource–performance comparison for open-weight LLM-based extraction in hydropower regulatory documentation and offers practical guidance for model selection under different deployment constraints.

Evaluation protocol

SYS-5620-Project - Achieving Robust Laser Performance utilizing Historical Shot Experiments - A Systems Engineering Proposal

The National Ignition Facility (NIF) at Lawrence Livermore National Laboratory precisely guides, amplifies, reflects, and focuses 192 powerful laser beams into a target about the size of a pencil eraser in a few billionths of a second, delivering more than 2 million joules of ultraviolet energy and 500 trillion watts of peak power. A crucial goal of the system is to trigger precise implosions of fuel capsules. This is achieved by delivering all 192 beams at user-specified times and locations on the target, minimizing any deviation from the requested performance. Power requirements can vary substantially on each experiment and the facility supports numerous amplifier pumping configurations and their attendant nonlinear effects. To achieve the tight performance required across such a broad array of configurations, constant comparison of measured and requested power delivery are tracked and long-term trends analyzed as a guide to understanding future performance. A very common question asked is given the current state of the laser today – how well would a similar experiment from the past perform today? Additionally, if one were to specify an alternate amplifier configuration using today’s model would and damage limits be exceeded and would there be an increase in performance. The physics model used to make these predictions and equipment protection checks is called the Virtual Beamline (VBL). VBL is used with an incoming desired pulse shape and energy to be delivered on target, and then does an iterative solve to predict the needed injected pulse in the front-end of the system to achieve this result. To effectively guide and predict future NIF experiment performance, laser scientists explore current amplifier configurations and compare them with historical data utilizing a tool called the Reverify Toolbox.

42 ENGINEERING

A Joint Search for Muon Neutrino Disappearance with the Short-Baseline Neutrino Program Using the SPINE Deep Learning-Based Reconstruction Package

We present the status of a joint search for muon neutrino disappearance in the Booster Neutrino Beam at Fermilab using the Short-Baseline Neutrino (SBN) Program's two-detector configuration, SBND and ICARUS. Charged-current interactions consistent with muon neutrinos and containing only a muon and at least one proton in the final state are reconstructed and selected using the SPINE deep learning-based particle reconstruction package. To exploit proton multiplicity information and enhance sensitivity to modeling effects, the selected sample is partitioned into exclusive channels with exactly one reconstructed proton and with more than one reconstructed proton. Comprehensive systematic uncertainties from the neutrino flux, interaction, and detector response are incorporated into the analysis, and the coverage of these systematic uncertainty models is validated using data from both detectors, including checks of near-far consistency in key kinematic and topology-sensitive observables. This analysis is intended for inclusion in SBN's first oscillation result.

Mueller, Justin [Fermilab]

Subject-specific modeling framework for particle deposition using computational fluid dynamics

Quantifying particle deposition and dose in the respiratory tract requires a physiologically realistic representation and reproducible computational workflows. However, existing modeling frameworks, such as the International Commission on Radiological Protection (ICRP) compartmental models and the Multiple Path Particle Dosimetry (MPPD) tool, lack detailed deposition profiles and subject-specific capabilities. The combination of advances in computer vision algorithms applied to the respiratory tract and Computational Fluid and Particle Dynamics (CFPD) allows high-fidelity simulations of particle behavior in anatomically accurate geometries derived from individual CT scans. The segmentation, preprocessing, and file preparation task for a CFPD simulation was often time-consuming, and no prior studies to-date have yet presented a fully automated framework. This work presents a fully automated workflow to obtain individualized particle deposition profiles in the human respiratory tract. The pipeline starts with segmenting upper and lower airway geometries using morphological and deep learning-based methods, generating three-dimensional (3D) models from CT imaging data. Next, a series of algorithms are presented to quality check and prepare the 3D geometry for a CFD or CFPD simulation. The preprocessing step includes correcting geometric artifacts, enforcing a physically consistent mesh, and automatically identifying and capping multiple outlets, which is required for CFD/CFPD simulations. These processed models are then input into open-source (OpenFOAM) or commercial (StarCCM+) CFD solvers, where flow and transient particle transport equations — including turbulence and particle–wall interactions are solved under realistic breathing conditions. Finally, the resulting particle deposition profiles can be integrated with Monte Carlo radiation transport codes and state-of-the-art computational phantoms to assess organ-specific absorbed doses in scenarios of radioactive aerosol inhalation. The presented work streamlines respiratory tract segmentation, preprocessing for CFD/CFPD simulations, and integration with dose assessment workflows, reducing manual intervention and improving access to high-fidelity, subject-specific modeling. The high precision in predicted particle deposition and dose distributions can improve personalized treatment strategies in respiratory medicine and refine dose estimates for radiation protection.

AI

LeWRON: Agentic Analysis of Electroweak Phase Transitions

The electroweak phase transition (EWPT) is a central topic in particle physics and cosmology, connecting collider phenomenology, baryogenesis, and gravitational-wave observatories. Its analysis requires a technically demanding, convention-sensitive, and model-dependent pipeline, from constructing the finite-temperature effective potential to tracking thermal histories, computing bubble nucleation rates, and predicting gravitational-wave spectra. We present LeWRON (Learning ElectroWeak phase tRansitiON), an agentic framework that orchestrates this pipeline starting from an input Lagrangian. LeWRON combines audited toolbox construction with an Explorer module that uses the generated model-specific code for further analysis, including scans and plots. Intermediate analytic outputs are checked by auditor agents and stored as structured artifacts, enabling reproducible human inspection and downstream use through both a command-line interface and a public Python API. The framework supports a reproduction mode, which infers conventions from the literature and reproduces published results, and a discovery mode, which guides users through structured checkpoints for new models. We demonstrate LeWRON across representative beyond-the-Standard-Model scenarios and release the code on GitHub.

Wang, Isaac R. [Fermilab] (ORCID:000000030789218X)

Conversational Grid Storage: Bridging Rucio and LLMs with Model Context Protocol

Experiments at Fermilab use Rucio to handle datasets that can be up to exabyte scale. However, navigating through Rucio’s syntax-heavy Command Line Interface (CLI) is a major workflow obstruction for researchers who just want to check quotas, track data identifiers (DIDs), or locate data sets. This project introduces a natural language interface. By building a containerized Model Context Protocol (MCP) server, an AI agent is created that translates plain English queries into data operations.

Akella, Kashyap [Fermilab; Illinois U., Urbana (ma

Development of an ERT‐Based Framework for Bentonite Buffers Monitoring From Laboratory Tests: 2. Quantitative Moisture Dynamics Estimation Model

Abstract The long‐term containment of high‐level radioactive waste in geological disposal repositories relies on Engineered Barrier Systems (EBS), with bentonite clay emerging as a candidate material due to its unique properties. Understanding moisture dynamics within bentonite buffers is crucial for EBS performance, as it directly influences the material's swelling capacity, thermal and hydraulic conductivity, mechanical properties, and long‐term evolution under complex thermal‐hydrological‐mechanical (THM) processes. This study develops an advanced Electrical Resistivity Tomography (ERT)‐based framework to quantitatively monitor moisture dynamics under THM conditions. Our framework extends the Waxman‐Smits model to incorporate the coupled effects of temperature, water content, fluid chemistry, and mechanical changes on bentonite's electrical properties. Utilizing HotBENT‐Lab data from our companion paper, which includes electrical conductivity, CT density, and thermocouple measurements, this study offers a novel methodological framework bridging different scales of the model. Our results show that the extended model can estimate water content from ERT data, capturing spatial and temporal variations in moisture distribution within bentonite columns. However, the model tends to overestimate water content compared to CT density‐derived measurements. We address this discrepancy by incorporating a simplified swelling effect model, which improves agreement between ERT and CT density‐based water content estimates. We also discuss model limitations, including simplified treatment of swelling and micropore effects, and propose a conceptual framework for transitioning from laboratory to field applications, addressing challenges such as parameter scalability, field validation methods, and integration of diverse data sources. This ERT‐based framework can potentially advance real‐world moisture monitoring of bentonite‐based EBS in nuclear waste repositories. Plain Language Summary Safely containing high‐level radioactive waste depends on barriers made from materials like bentonite clay, which is effective because it swells and seals in the waste. To ensure these barriers work well over time, it's important to understand how moisture moves through the clay. Our study developed a new method using ERT to monitor moisture levels in bentonite under conditions that mimic those in actual storage sites, including changes in temperature, water content, and mechanical stress. This study improved an existing model to better account for how these factors affect the clay, allowing us to create more accurate moisture maps. Initially, the proposed model overestimated the amount of water in the clay, but its accuracy was improved by factoring in how the clay swells when wet. This study also identified some limitations of the model and suggested ways to adapt it for use in real‐world waste storage sites. This new approach could lead to better monitoring and safety checks for nuclear waste storage systems, helping to ensure long‐term containment. Key Points This work develops an ERT‐based framework extending the Waxman‐Smits model to monitor bentonite moisture dynamics during coupled THM processes The extended model accurately estimates water content from Electrical Resistivity Tomography data, incorporating swelling effects to improve precision This work proposes a conceptual framework for transitioning from laboratory to field applications, advancing EBS monitoring in nuclear waste repositories

Chen, Hang

An On-Sky Atmospheric Calibration of SPT-SLIM

We present the methodology and results of the on-sky responsivity calibration of the South Pole Telescope Shirokoff Line Intensity Mapper (SPT-SLIM). SPT-SLIM is a pathfinder line intensity mapping experiment utilizing the on-chip spectrometer technology, and was first deployed during the 2024-2025 Austral Summer season on the South Pole Telescope. During the two-week on-sky operation of SPT-SLIM, we performed periodic measurements of the detector response as a function of the telescope elevation angle. Combining these data with atmospheric opacity measurements from an on-site atmospheric tipping radiometer, simulated South Pole atmospheric spectra, and measured detector spectral responses, we construct estimates for the responsivity of SPT-SLIM detectors to sky loading. We then use this model to calibrate observations of the moon taken by SPT-SLIM, cross-checking the result against the known brightness temperature of the Moon as a function of its phase.

Dibert, K. R. [Chicago U., Astron. Astrophys. Ctr.

Numerical Modeling & Optimization of the iProTech Pitching Inertial Pump (PIP) Wave Energy Converter (WEC) (CRADA Final Report)

This project represents a continuation of the collaboration between iProTech and NLR to simulate, optimize and design the iProTech Pitching Inertial Pump (PIP) device. The objectives of this TEAMER project are twofold: 1. Refining the physical characteristics of the existing iProTech PIP WEC-Sim model to enhance the model’s fidelity and include controllable components. Key model enhancements target the inclusion of Coulomb friction, the introduction of a controllable bypass valve, and the replacement of traditional check valves with advanced motorized ones. 2. Exploring traditional and advanced control algorithms. From traditional methods like latching control to cutting-edge reinforcement learning (RL) algorithms, the goal is to ensure the PIP device's adaptability and optimal performance across a range of ocean conditions. NLR is tasked with augmenting the WEC-Sim model and implementing the control algorithms, culminating in performance comparison analyses. iProTech will update their existing 3D models, advise on model improvements, and determine crucial system metrics. WEC-Sim, developed in MATLAB/SIMULINK with Simscape Multibody, is the main piece of software that will be used in this project. Coupled with the MATLAB RL Toolbox, it offers a robust platform for in-depth simulation and optimization of the iProTech PIP device. Building on previous work to explore the PIP design space and optimize its geometry, mass distribution, center of gravity and other key parameters, this project aims to refine iProTech’s existing numerical models and develop effective control algorithms that can seamlessly integrate into their future hardware testing campaigns.

16 TIDAL AND WAVE POWER

FY25 Theory and Simulation Performance Target: Development of an integrated modeling framework for fusion reactor design and assessment (Final Report)

This report documents the FY25 Theory and Simulation Performance Target (TSPT) of developing an integrated modeling framework for fusion reactor design and assessment (FREDA). Over Q1-Q4, new capabilities were developed across both plasma and engineering domains and demonstrated on an example representation of a Compact Advanced Tokamak with a Dual Cooled Lead Lithium blanket. This represents a first-of-a-kind demonstration of coupled core-to-wall-to-engineering for a reactor. Self-consistent CESOL workflows were applied to provide core, pedestal, and SOL prediction; new modules were developed for energetic particle stability (FAR3D) and transport (TGLF-EP) analysis; and boundary plasma modeling (SOLPS-ITER, BOUT++/Hermes-3) was expanded to evaluate wall and divertor heat fluxes and interface with engineering thermal analysis. A parameterized CAD tool, TRACER, was expanded to generate medium-fidelity divertor, blanket, and coil geometries; OpenFOAM and Diablo workflows were applied for first-wall and divertor thermal analyses with helium cooling; and reduced-order models were created for high-mass-flux divertor cooling. Magnet multiphysics capabilities were verified between Elmer, Diablo, and a new MFEM-based solver, and workflows enable stress, thermal, and neutron-fluence analysis of TF coils with neutronics-driven heating. Nuclear and blanket analysis workflows were demonstrated, including tritium breeding, transport, and CFD-informed thermo-mechanical assessment. Preliminary multi-fidelity uncertainty quantification workflows were applied to boundary modeling codes and shown to achieve variance reductions with fewer high-fidelity boundary simulations. Key findings highlight the challenges of resolving the ITEP gap to find suitable balance between wall and divertor loads, neutron heating, and practical limits of PFC cooling. Next step priorities are to develop automated workflows to check boundary code convergence and detachment, implement tighter physics-engineering CAD provenance tracking, and inclusion of plasma-material interface models for SLAG and tungsten cracking behavior. Collectively, these developments establish sophisticated capabilities for predictive, multi-fidelity, whole-device modeling that integrates plasma physics, materials, magnets, and nuclear engineering to guide pathways to viable Fusion Pilot Plant design points.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY

Arm and shoulder muscle segmentation in axial MRI with UNet deep learning model

Quantifying individual upper-limb muscle volumes from MRI provides key insight into muscle-specific strength, deficits, and adaptations. Manual delineation is the gold standard but time‑intensive, and the performance of current deep learning approaches, particularly for small or anatomically complex muscles, remains incompletely characterized. We evaluated a state‑of‑the‑art deep learning framework across the entire upper limb and analyzed factors governing segmentation performance, with attention to the forearm. Three previously published MRI datasets (1.5 T, 3D GRE T1‑weighted; total n = 39) spanning young, middle‑aged, and older adults were curated and quality‑checked, including expert manual segmentations for 31 muscles. Following multiclass mask reconstruction, we trained three 3D nnU‑Net multiclass models matched to the muscle subsets present across datasets, using five‑fold cross‑validation and a composite Dice Similarity Coefficient (DSC) + cross entropy loss. Segmentation accuracy was assessed with DSC. Performance varied across muscles (mean DSC = 0.806 ± 0.098), ranging from 0.920 (Deltoid) to 0.461 (Extensor pollicis brevis). In uncertainty‑weighted regressions, muscle volume was positively associated with DSC (R2 = 0.36, p < 0.001), whereas training segmentation count and muscle orientation showed negligible associations (R2 ≤ 0.06). A weighted mixed‑effects model identified volume as the strongest evaluated predictor, explaining 23.9% of variance in DSC; orientation and training count each contributed <1%, leaving 61.5% unexplained. These results indicate that deep learning–based segmentation can accurately quantify muscle volume for many upper‑limb muscles but remains constrained for small, low‑contrast forearm muscles.

Gillespie, Samuel

Cosmological constraints from the cross-correlation of DESI Luminous Red Galaxies with CMB lensing from Planck PR4 and ACT DR6

Here, we infer the growth of large scale structure over the redshift range 0.4 ≲ z ≲ 1 from the cross-correlation of spectroscopically calibrated Luminous Red Galaxies (LRGs) selected from the Dark Energy Spectroscopic Instrument (DESI) legacy imaging survey with CMB lensing maps reconstructed from the latest Planck and ACT data. We adopt a hybrid effective field theory (HEFT) model that robustly regulates the cosmological information obtainable from smaller scales, such that our cosmological constraints are reliably derived from the (predominantly) linear regime. We perform an extensive set of bandpower- and parameter-level systematics checks to ensure the robustness of our results and to characterize the uniformity of the LRG sample. We demonstrate that our results are stable to a wide range of modeling assumptions, finding excellent agreement with a linear theory analysis performed on a restricted range of scales. From a tomographic analysis of the four LRG photometric redshift bins we find that the rate of structure growth is consistent with ΛCDM with an overall amplitude that is ≃ 5-7% lower than predicted by primary CMB measurements with modest (∼ 2σ) statistical significance. From the combined analysis of all four bins and their cross-correlations with Planck we obtain S 8 = 0.765 ± 0.023, which is less discrepant with primary CMB measurements than previous DESI LRG cross Planck CMB lensing results. From the cross-correlation with ACT we obtain S 8 = 0.790 +0.024 -0.027 , while when jointly analyzing Planck and ACT we find S 8 = 0.775 +0.019 -0.022 from our data alone and σ 8 = 0.772 +0.020 -0.023 with the addition of BAO data. These constraints are consistent with the latest Planck primary CMB analyses at the ≃ 1.6-2.2σ level, and are in excellent agreement with galaxy lensing surveys.

cosmological parameters from LSS

SCALE Procedure for Verified, Archived, Library of Inputs and Data (VALID)

This procedure provides a framework for preparing, reviewing, and storing model inputs and derived data so that individuals with authorized access to the Verified, Archived, Library of Inputs and Data (VALID) repository can use the inputs and data with confidence in their analyses. This procedure uses documented checks and reviews to ensure that the inputs and data were correctly generated using appropriate references. Configuration management is implemented to prevent inadvertent modification of the inputs and data or inclusion of models that have not been reviewed. This procedure also provides guidance to be followed if errors are identified or if input or data revisions are needed.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS

Numerical challenges for energy conservation in N -body simulations of collapsing self-interacting dark matter halos

Dark matter (DM) halos can be subject to gravothermal collapse if the DM is not collisionless, but engaged in strong self-interactions instead. When the scattering is able to efficiently transfer heat from the centre to the outskirts, the central region of the halo collapses and reaches densities much higher than those for collisionless DM. This phenomenon is potentially observable in studies of strong lensing. Current theoretical efforts are motivated by observations of surprisingly dense substructures. However, a comparison with observations requires accurate predictions. One method to obtain such predictions is to use N-body simulations. Collapsed halos are extreme systems that pose severe challenges when applying state-of-the-art codes to model self-interacting dark matter (SIDM). In this work, we investigate the root of such problems, with a focus on energy non-conservation. Moreover, we discuss possible strategies to avoid them. We ran N-body simulations, both with and without SIDM, of an isolated DM-only halo and we adjusted the numerical parameters to check the accuracy of the simulation. We find that not only the numerical scheme for SIDM can lead to energy non-conservation, but also the modelling of gravitational interaction and the time integration are problematic. The main issues we find are: (a) particles changing their time step in a non-time-reversible manner; (b) the asymmetry in the tree-based gravitational force evaluation; and (c) SIDM velocity kicks breaking the time symmetry. Tuning the parameters of the simulation to achieve a high level of accuracy allows us to conserve energy not only at early stages of the evolution, but also later on. However, the cost of the simulations becomes prohibitively large as a result. Some of the problems that make the simulations of the gravothermal collapse phase inaccurate can be overcome by choosing appropriate numerical schemes. However, other issues still pose a challenge. Our findings motivate further works on addressing the challenges in simulating strong DM self-interactions.

dark matter