Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “open-source tools”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Time-Domain Vortex Induced Vibration Modeling of Reference Dynamic Power Cable for the Gulf of Maine

Vortex-induced vibration (VIV) is a phenomenon known to decrease the fatigue life of dynamic power cables used in floating offshore wind systems through increased bending loading cycles. A variety of modeling solutions have been proposed to study this, but none enable fully-coupled simulations leveraging open-source tools. To address this, a time-domain VIV model has been successfully implemented into the open-source mooring dynamics model MoorDyn. The semi-empirical VIV model describes a lift force acting on segments of a flexible cylinder, making it well suited for implementation into MoorDyn. The new capability successfully predicted the frequencies and magnitudes of strain in both steady and oscillating flows when compared to the original validation results and matched peak spectral response when compared to experimental results, verifying successful implementation. MoorDyn with VIV was then leveraged to simulate a 15 MW floating turbine with a dynamic power cable in the Gulf of Maine using OpenFAST for six return periods of combined wind, wave, and current conditions. Increased curvature and tension fluctuations in higher flow speeds were observed when simulating VIV. Maximum tensions and curvatures also increased, with the 500 year conditions violating the curvature factor of safety of 2.0. These results highlight the importance of considering VIV when designing dynamic power cables for the Gulf of Maine, and indicate that cost effective mitigation strategies should be explored for the area. They also demonstrate the utility of this new modeling capability, which provides the first open-source tool for fully coupled time-domain floating offshore wind simulations with dynamic power cable VIV.

17 WIND ENERGY↗

3D Play Fairway Analysis for Examining of Superhot Reservoir Production Scenarios

The DEEPEN (DE-risking Exploration for geothermal Plays in magmatic ENvironments) project was a multi-laboratory, international effort to reduce uncertainty and improve resource characterization in superhot geothermal systems. Building on this foundation, this work advances open-source tools designed to lower the exploration risk and cost of superhot geothermal projects while promoting transparency, reproducibility, and efficiency in exploration workflows. These tools are being tested at two key sites: (1) the Nesjavellir Geothermal Area in Iceland, where the Icelandic Deep Drilling Project (IDDP) will drill its third well, and (2) Newberry Volcano in Oregon, USA, where Mazama Energy will pilot the first superhot enhanced geothermal system (EGS). A major outcome is the creation of a modular, open-source Python framework for play fairway analysis (PFA) in 2D and 3D, called geoPFA. The PFA workflow has been expanded to produce pseudo conceptual models, and will soon be refined to assess reservoir components through integration with the thermo-hydraulic-mechanical-chemical (THMC) simulator TReactMech, to enable iterative coupling between PFA and THMC models, improving characterization of superhot systems. All three of the Icelandic Deep Drilling Project's production scenarios were analyzed via this framework: (1) a superhot deep injection well paired with conventional production wells at Nesjavellir, (2) a superhot deep production well at Nesjavellir, and (3) superhot enhanced geothermal system at Newberry Volcano. This analysis provides useful insights around conceptual modeling of these production scenarios, helping to inform decisions around which scenario is best suited for which types of environments.

15 GEOTHERMAL ENERGY↗

Future Projections of Lifecycle Cost and Greenhouse Gas Emissions of Light-Duty Vehicles

Vehicles with electrified powertrains carry the promise of significant reductions in greenhouse gas (GHG) emissions from a lifecycle analysis (LCA) standpoint compared to conventional internal combustion engine (CICE) vehicles. However, trade-offs exist between different types of electrified powertrains in terms of cost, consumer acceptance, and GHG reduction efficacy for different operating conditions. The open-source tool CarGHG was developed with an aim to enable the exploration of a plethora of parametric study scenarios, including the cost of electrification technologies, different driving patterns and charging habits, and the cost and carbon intensity of electricity and fuel blends. This paper introduces the framework of CarGHG, then showcases total cost of ownership (TCO) and LCA GHG results for select models of light-duty vehicles. Another capability of CarGHG, which is the ability to estimate the performance of “virtual” vehicle models (perceived vehicle design specifications not yet on the market), is utilized to explore future scenarios of electrification and low-carbon fuel blends for Small Sports Utility Vehicles (SUVs), a popular light-duty vehicle segment in North America. With opportunities, but also uncertainties, in future scenarios, it is likely wise to continue pursuing multiple ways towards the reduction of LCA GHG.

Hamza, Karim↗

Simulating Thermoelectric Devices Using the MOOSE Framework

Thermoelectric generators (TEG) are devices that generate energy by converting heat into electricity or provide cooling via the Peltier effect. This feature of thermoelectric devices originates from the Seebeck, Peltier, Thomson, and Joule heating effects. TEGs can be applied in energy and thermal management systems such as waste heat recovery and refrigeration, respectively. Thermoelectric device design is influenced by the material selection and the device's geometry operating conditions. Therefore, predicting, verifying, and validating thermoelectric device performance using simulations tools is essential to deploying thermoelectric devices in industry. The Multiphysics Object-Oriented Simulation Environment (MOOSE) Framework is an open-source simulation tool capable of modeling simple to complex systems. In this work, we demonstrate MOOSE's thermoelectric device modeling capabilities by simulating a unicouple, module, and exhaust gas recovery system. The Seebeck, Peltier, Thomson, and Joule heating physics are implemented into MOOSE. The MOOSE thermoelectric physics were thoroughly verified and validated using published COMSOL® results and experimental data. In addition, thermoelectric modules were integrated into an exhaust gas recovery system using the MOOSE MultiApp function as a demonstration of the model's ability. The verification and validation results and exhaust gas heat recovery system showcases MOOSE's capability to model thermoelectric devices and integrate these devices into practical energy systems.

42 - ENGINEERING↗

ML-Shock-Time-Series-Synthesis

Open-source machine learning tools for GPU-batched synthetic shock time-series generation, GPU-accelerated batched Shock Response Spectrum (SRS) computation, and standardized benchmark datasets.

Watts, Adam↗

Web-based Preprocessing and Visualization of 3D FIB Tomography Data for Nuclear Fuel Characterization

Three-dimensional (3D) focused ion beam (FIB) tomography enables reconstruction of internal nuclear fuel features that can't be fully evaluated through surface imaging alone. This capability supports characterization of fuel constituents and defects under thermal and irradiation conditions relevant to microreactor development. However, large tomography datasets can create data-handling, loading, and visualization challenges, especially when image-stack preparation and file conversion must be completed with separate tools. The Computational Ultraspatial Tomography Toolkit for High-Resolution Object Analysis Tools (CUTTRHOAT) is an open-source web application being developed to display FIB tomography datasets available through the Nuclear Research Data System (NRDS). The current alpha version requires prepared HDF5 datasets and has limited integrated data-preparation capabilities. This project improves CUTTHROAT by adding dataset-folder selection, automatic input detection, dataset scanning, missing-slice identification, blank-slice insertion, and image-stack-to-HDF5 conversion. Two applications will be compared: the baseline CUTTHROAT alpha workflow and the updated application containing the integrated data-handling and preprocessing functions. Evaluation will consider dataset detection accuracy, conversion success, loading time, rendering responsiveness, application stability, and user interaction. Preliminary results demonstrate successful loading of existing HDF5 files and converted image stacks, while testing also identified performance reductions caused by excessive blank-slice generation. The updated workflow reduces reliance on external preparation tools and supports more direct movement from image stacks to color-code 3D visualization. Future work includes refining missing-slice handling, integrating additional preprocessing functions, like a denoising feature, parsing TIFF metadata for automatic voxel scaling, and adding manual X, Y, and Z voxel-spacing inputs for PNG and JPEG.

36 - MATERIALS SCIENCE↗

Validation of OpenPronghorn for Periodic Hill Flow Separation

OpenPronghorn is an open-source, MOOSE-based thermal-hydraulics simulation tool used for advanced reactor analysis. As an open-source code, it offers a transparent framework for validating governing equations, assumptions, and numerical methods against established benchmarks. This study evaluates OpenPronghorn's RANS turbulence model against the ERCOFTAC Case 81 periodic hill benchmark, a standard test case for separated turbulent flow featuring curved-wall separation, recirculation, shear-layer development, and reattachment. Streamwise velocity profiles predicted by OpenPronghorn were compared to reference LES data at multiple x/h locations. Results show that OpenPronghorn captures the overall trend of the velocity profiles, but the largest discrepancies occur in the separated-flow region, where turbulence is highly anisotropic and strongly affected by adverse pressure gradients and wall curvature—conditions that are inherently difficult for standard RANS models to resolve. Future work will test alternative k-e model variants and correction terms to improve prediction accuracy in this region.

42 - ENGINEERING↗

Impact of Limited Degree of Freedom Drag Coefficients on a Floating Offshore Wind Turbine Simulation

The worldwide effort to design and commission floating offshore wind turbines (FOWT) is motivating the need for reliable numerical models that adequately represent their physical behavior under realistic sea states. However, properly representing the hydrodynamic quadratic damping for FOWT remains uncertain, because of its dependency on the choice of drag coefficients (dimensionless or not). It is hypothesized that the limited degree of freedom (DoF) drag coefficient formulation that uses only translational drag coefficients causes mischaracterization of the rotational DoF drag, leading to underestimation of FOWT global loads, such as tower base fore-aft shear. To address these hydrodynamic modeling uncertainties, different quadratic drag models implemented in the open-source mid-fidelity simulation tool, OpenFAST, were investigated and compared with the experimental data from the Offshore Code Comparison Collaboration, Continued, with Correlation (OC5) project. The tower base fore-aft shear and up-wave mooring line tension were compared under an irregular wave loading condition to demonstrate the effects of the different damping models. Two types of hydrodynamic quadratic drag formulations were considered: (1) member-based dimensionless drag coefficients applied only at the translational DoF (namely limited-DoF drag model) and (2) quadratic drag matrix model (in dimensional form). Based on the results, the former consistently underestimated the 95th percentile peak loads and spectral responses when compared to the OC5 experimental data. In contrast, the drag matrix models reduced errors in estimates of the tower base shear peak load by 7–10% compared to the limited-DoF drag model. The underestimation in the tower base fore-aft shear was thus inferred be related to mischaracterization of the rotational pitch drag and the heave motion/drag by the limited-DoF model.

17 WIND ENERGY↗

CHEQUP v0.1

CHEQUP (Castro-based Hofi Expansion with QUasineutral Plasma) is a simulation code for modeling the formation of hydrodynamic optical-field-ionized (HOFI) plasma channels, which are used as waveguides in laser-plasma acceleration experiments. This includes experiments performed at LBNL's BELLA facility as well as other laser facilities across the world. CHEQUP extends the open-source Castro hydrodynamics framework with physics modules tailored for modeling HOFI plasma channels -- including multi-species ionization and three-body recombination for mixtures of hydrogen, nitrogen, helium, and argon ; a two-temperature model tracking electron and heavy-species temperatures separately ; and coupling with other codes of the BLAST ecosystem (https://blast.lbl.gov/) such as WarpX, via the openPMD standard. CHEQUP inherits from Castro the ability to run on modern GPU architectures (NVIDIA CUDA, AMD HIP) and supports adaptive mesh refinement (AMR) for efficient multi-scale resolution. Compared to existing tools, CHEQUP would be, to our knowledge, the first open-source code implementing the full HOFI channel formation physics, and the first implementation capable of running on GPUs. This enables significantly faster, large-scale parameter scans critical for the design of next-generation LPA-based accelerators and light sources.

Lehe, Remi [Lawrence Berkeley National Laboratory ↗

Generalized Tensor-on-Tensor Regression (GToTR)

SAND2026-23069O Generalized Tensor-on-Tensor Regression (GToTR) is a Python-based tool for conducting generalized tensor-on-tensor regression. It provides Canonical Polyadic (CP)-based generalized tensor regression models, support for generalized linear model-like families and links, alternating-optimization model fitting methods, and a standard statistics software interface. The tool supports tensor-valued responses and covariates using the open-source Python Tensor Toolbox (pyttb) software package. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Dunlavy, Daniel [Sandia National Lab. (SNL-CA), Li↗

DTLMod: A simulation framework for in situ workflow optimization

In situ processing workflows have become essential for coping with the explosion in data volume and velocity in large-scale scientific computing, providing domain scientists with early insights at runtime. Multiple frameworks implement this paradigm through a data transport layer (DTL), offering different data access modes and deployment schemes, but researchers currently lack the appropriate tools to assess design and deployment options before committing to costly real experiments. We introduce DTLMod, an open-source simulated DTL that enables performance evaluation of in situ workflow configurations at scale. Built on SimGrid, it links into any SimGrid-based simulator and is available in C++ and Python. We evaluate DTLMod along four axes: scalability (tens of thousands of simulated processes across interconnected clusters in seconds, with linear memory scaling), versatility (three implementation variants trading fidelity for speed), accuracy (simulated times faithfully reflecting real behavior), and practical utility (two use cases demonstrating evidence-based workflow design decisions).

Suter, Fred [ORNL] (ORCID:0000000319021955)↗

SetGo: Metadata Readiness for Scientific AI Datasets

Scientific datasets intended for AI use require both computational readiness for model training and metadata readiness for discovery, sharing, and reuse. The Readiness Engine for Data Integration (REDI) addresses computational readiness, but no corresponding tool evaluates whether a dataset’s metadata are sufficiently complete, governed, and standards-compliant for publication and agent-based consumption. Existing FAIR assessors operate only on published repository records, and no single system covers FAIR compliance, licensing, provenance, governance, reproducibility, and catalog readiness together. We present SetGo, an open-source Python toolkit that assesses and repairs metadata readiness across these six dimensions before a dataset is published or archived. Applied to four scientific corpora, SetGo surfaces deficiencies that general-purpose tools do not detect: ERA5 climate metadata scores 4% on ACDD 1.3 compliance; materials datasets fail OPTIMADE species-definition requirements; and PDB-derived proteomics data carries licensing terms incompatible with standard SPDX identifiers. Guided enrichment raises overall FAIR scores from 52–57% to 81–91%, and a single setgo publish command pushes to Hugging Face Hub, CKAN, or OpenMetadata with ML Commons Croissant 1.0 metadata sidecars. To support interactive and automated workflows, SetGo integrates with coding agents powered by large language models (LLMs) through a /setgo skill that enables natural-language execution of the full assess–enrich–publish loop, with user involvement limited to supplying missing metadata values.

Wilkinson, Sean [ORNL] (ORCID:0000000214437479)↗

From natural language to control signals: a conceptual framework for semantic channel finding in complex experimental infrastructure

Modern experimental platforms such as particle accelerators, fusion devices, telescopes, and industrial process control systems expose tens to hundreds of thousands of control and diagnostic channels, accumulated over decades of hardware evolution. Operators and AI systems alike depend on informal expert knowledge, inconsistent naming conventions, and scattered documentation to locate the signals required for monitoring, troubleshooting, and automated control, creating a persistent bottleneck for reliability, scalability, and emerging language-model-driven interfaces. We formalize semantic channel finding, the task of mapping natural-language intent to concrete control-system signals, as a general problem in complex experimental infrastructure, and introduce a four-paradigm conceptual framework to guide architecture selection based on facility-specific data regimes. The paradigms span (i) direct in-context lookup over small, curated channel dictionaries, (ii) constrained hierarchical navigation through structured trees, (iii) interactive agent exploration using iterative reasoning and tool-based database queries, and (iv) ontology-grounded semantic search that decouples channel meaning from facility-specific naming conventions. We demonstrate the practical feasibility of each paradigm through proof-of-concept implementations at four operational facilities spanning two orders of magnitude in scale: from compact free-electron lasers to large synchrotron light sources, operating under diverse control-system architectures ranging from clean hierarchical naming schemes to legacy environments with decades of heterogeneous conventions. Where evaluated against expert-curated operational queries, these instantiations achieve 90%–97% accuracy, validating the framework’s applicability across real-world deployment scenarios. To accelerate adoption across the broader scientific and industrial control-system community, we release open-source, plug-and-play implementations of all three interactive paradigms-direct lookup, hierarchical navigation, and middle-layer exploration-within the Osprey framework, together with tools for channel database generation, interactive testing, and minimal-configuration deployment. This work establishes semantic channel finding as a foundational capability for human-centric and agentic AI interfaces at large-scale facilities, providing both a systematic framework for architecture design and practical resources to enable adoption without building custom infrastructure from scratch.

channel finding↗

Self-Admitted Technical Debt in Scientific Software: Prioritization, Sentiment, and Propagation Across Artifacts

Self-admitted technical debt (SATD) impairs scientific software (SSW), yet its prioritization, sentiment, persistence, and propagation remains underexplored. Understanding how SSW developers express, and address SATD is crucial for improving SSW maintenance, and tooling. This study investigates how SATD types and artifacts in SSW are prioritized, how sentiment relates to urgency, SATD removal and resolution rates, and the extent to which SATD propagates across artifacts. We analyzed nine SSW repositories using a SATD classification model and a semantic embedding-based prioritization heuristic. SATD was examined across multiple artifacts, with sentiment assessed via a fine-tuned transformer. Propagation was traced, priority scores compared to static analysis, and removal and resolution rates quantified. SATD in comments, commits, and pull requests receive higher priority than SATD in issues, with negative sentiment amplifying urgency. Resolution and removal rates lag behind open-source software (OSS) averages. Most SATD remains confined to the originating artifact, but longer propagation chains are rare and correlate with higher priority, highlighting persistent and high impact debt. Prioritization is influenced by artifact type and sentiment, while low removal and resolution rates signal persistent debt. Cross-artifact propagation marks high priority, unresolved SATD, providing empirical guidance for targeted monitoring, review prioritization, and tool supported maintenance in SSW.

Melin, Eric [Boise State University]↗

Vision and Development of a Design, Implementation, and Verification Automation (DIVA) Software Platform for DNA Construction

Abstract DNA construction, while a prerequisite to many biological endeavors, is often a time-consuming distraction from an individual’s primary research objectives. We envisioned that with the right software infrastructure and cultural mindset, a single person could execute in parallel the batched DNA construction tasks of an entire research institute, at scales realizing efficiency gains through process and laboratory automation. In pursuit of this vision, we developed the Design, Implementation, and Verification Automation (DIVA) software platform. DIVA’s web interface enables researchers to design DNA constructs (using visual biological computer-aided design tools and biological parts repositories), submit designs for construction to dedicated staff, and track DNA construction as it progresses. DIVA supports the dedicated staff through the DNA construction process and records both successful and unsuccessful attempts toward improving the overall process. The platform is publicly available at public-diva.jbei.org and its open-source code through github.com/JBEI/DIVA.

Plahar, Hector [DOE Agile BioFoundry , , ,; DOE Jo↗

Powered By SAM [Slides]

The System Advisor Model(TM) (SAM) is a free, open-source desktop application for techno-economic analysis of energy technologies. By combining detailed performance modeling with financial analysis, SAM allows users to assess technology trade-offs, explore future scenarios, and make informed decisions about energy investments. Users also have access to model details and the ability to embed SAM's core models in their own applications. This webinar, hosted by National Laboratory of the Rockies researchers Janine Keith and Matt Prilliman, highlights how this widely used modeling tool supports data-driven decision-making for energy systems.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

A unified large language model–based framework for heterogeneous PV image diagnosis

With advances in imaging technologies, modern photovoltaic (PV) systems generate large volumes of heterogeneous image data, including visible, electroluminescence (EL), and infrared (IR) images. Existing PV image analysis models, particularly deep learning approaches, are typically task-specific and lack cross-modality generalization. To address this limitation, this paper proposes an open-source large language model (LLM)–based unified framework for heterogeneous PV image diagnostics. Through task-aware diagnostic prompting, the framework enables analysis of visible, EL, and IR images within a single pipeline, supporting both zero-shot and few-shot inference and binary and multiclass classification. It is compatible with state-of-the-art multimodal LLMs, including ChatGPT, Gemini, Claude, Qwen, and CLIP. The framework is evaluated on PV module condition classification (clean, soiling, snow, hail, and bird droppings) using visible images, cell crack detection using EL images, and hotspot detection using IR images. GPT-5.1 in few-shot mode achieves the best performance, with classification accuracy exceeding 97.3%. Open-source models such as Qwen and CLIP also deliver competitive results on visible images (around 90% accuracy), though their performance is more limited on EL and IR modalities. On the full ELPV dataset, the framework achieves 83.5% zero-shot accuracy, within 2.8% of the supervised CNN baseline, confirming scalability to larger benchmarks. Practical aspects such as reproducibility, response latency, and confidence estimation are systematically analyzed. The framework operates across PV image modalities without modality- or task-specific training, making it well suited as a rapid pre-screening tool to support downstream detailed diagnostics. A benchmark dataset of diverse labeled PV images is also released.

Li, Baojie↗