Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “DeepLynx”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

DeepLynx Ecosystem 2025

Poor data integration and governance continue to plague complex engineering projects, resulting in missed cost, schedule, and performance targets. Departments operate in isolated systems with manual data exchange, creating fragmented information that compounds errors and leads to significant delays and cost overruns. The DeepLynx ecosystem addresses these challenges through an open-source, modular data management platform that transforms fragmented project data into an integrated digital thread. Built on a federated microservice architecture, the ecosystem comprises seven specialized tools centered around DeepLynx Nexus, a unified data catalog with hierarchical organization and graph-based navigation capabilities. The ecosystem includes: DeepLynx Stream for real-time timeseries data ingestion from industrial sources; DeepLynx Ingest for governed data uploads with formal review workflows; DeepLynx Lattice for ontology-based entity and relationship extraction; DeepLynx Run for workflow orchestration and secure AI/ML compute; DeepLynx Visualize for 3D digital twin visualization; and DeepLynx Insight for AI-assisted document analysis with traceable, grounded responses. Deployable in cloud, on-premise, or hybrid environments using containerized Docker applications and Helm charts, the DeepLynx ecosystem provides flexible infrastructure that adapts to organizational requirements. By consolidating project data into a unified data lake with role-based access controls and OAuth2 authentication, DeepLynx enables digital thread and digital twin capabilities that improve decision-making, reduce risk, and support complex engineering workflows throughout the project lifecycle.

42 - ENGINEERING↗

Deeplynx Airflow Provider Package

The DeepLynx Airflow Provider Package is a python package used to interact with the data warehouse DeepLynx when using the workflow orchestration tool Apache Airflow. This python package is packaged together using the airflow package standard so that it can be easily installed and used in any Apache Airflow environment. This package is meant to encapsulate the DeepLynx API for use in Airflow so that any interactions with DeepLynx that a user may want to use in their Airflow workflow can be easily accomplished using this provider package. This allows us to develop, implement, and test our DeepLynx-Airflow interactions in one provider package repository, and then easily install and use this package in any airflow instance. This DeepLynx Airflow Provider Package will be used extensively by the DeepLynx DAG repository.

Cavaluzzi, JackM↗

Deeplynx Dag Repository

The DeepLynx DAG repository will contain several Airflow DAGs (Directed Acyclic Graphs) which will be used in the context of DeepLynx's deployed Apache Airflow instance. These DAGs will be used for multiple data management tasks for DeepLynx data, including but not limited to: - bringing data from various sources and tools into DeepLynx - managing sequential data workflows, such as running Python scripts on data to perform analysis and returning the results to DeepLynx - performing any necessary transformation or pre-processing on data coming into DeepLynx from external sources or out of DeepLynx to go to external applications

Brownlee, JarenM.↗

Improving the User Interface of the DeepLynx Data Warehouse

DeepLynx is an open-source ontology-based data warehouse created by INL to support the creation and life cycle of digital engineering projects, with a particular emphasis on digital twins [1]. Digital twins are systems that represent physical assets and process in a real-time digital environment [1]. Most well-known commercial data warehouses use Graphical User Interfaces (GUIs) for users to interact with their systems [3]. Limited publications have addressed the design of these interfaces and understanding of their target users. The current users and development team acknowledge the need to improve the current UI, not just for aesthetics but to improve functionality and workflow of DeepLynx. Traditional data warehouse users are developers, data scientists and business analysts [2]. DeepLynx users have a vast range of experience using data warehouses, and diverse roles, including engineers, scientists and management positions. Because there is a broader audience of target users for DeepLynx than a typical data warehouse, it is essential that DeepLynx has a useable and intuitive user interface. To achieve this the team performed human-computer interaction methods, including a Heuristic Evaluation of current UI using Neilsen’s Usability Heuristic, create personas based on current users by designing a user survey, data analysis and develop of personas. Followed by a redesign of the UI following using Neilsen’s Usability Heuristic and Norman’s Principles of Interactive Design in industry standard software Figma. Lastly a Heuristic Evaluation of new UI design, using Neilsen’s Usability Heuristic and User testing of redesign UI and have a group of users complete a Thinking Aloud Test of the new UI. Preliminary results of the Heuristic Evaluation of current UI arise issue with Consistency and Standards, Visibility of System Status, Match System and Real World and Recognition Rather than Recall. These issues were addressed in the proposed redesign by applying Neilsen’s Usability Heuristic and Norman’s Principles of Interactive Design. Next steps include formalized list of lessons learned and design implications for future publications.

97 MATHEMATICS AND COMPUTING↗

Deeplynx Supervisory Control Adapter

This software is intended to facilitate the sending of data from DeepLynx to some human machine interface (HMI). The HMI would read the data generated by this software and potentially make a physical change on a process which the HMI controls. This software enables a digital twin using DeepLynx to talk to an HMI, thereby making autonomous updates to the physical asset which is being twinned.

Wilsdon, KatherineN [Idaho National Laboratory] (0↗

Deeplynx-loader

'deeplynx-timeseries-loader' is a library designed to make it as easy as possible for users to download and access timeseries or tabular data from DeepLynx.

Darrington, John↗

Deeplynx Rust Sdk

This software is a Rust package that interacts with the Application Programming Interface (API) suite provided by DeepLynx. A Rust codebase may import this package in order to have access to these methods for communicating with a DeepLynx instance.

Browning, JerenM [Idaho National Laboratory (INL),↗

The DeepLynx Data Warehouse

Digital Engineering and the development of digital twins necessitate the use of highly sophisticated tools and software. These methodologies and systems bring many benefits but rely on accomplishing challenging goals such as disparate data and systems integration into a single cohesive and holistic system. Idaho National Laboratory has recently released an open-source data warehouse – DeepLynx – to help in the planning, creation, execution, and management of projects in the digital engineering space, with an emphasis on supporting digital twins in an operational space.

97 MATHEMATICS AND COMPUTING↗

P6 Deeplynx Adapter

The purpose of this software is to pull data from the Oracle Primavera P6 scheduling tool and bring it into the DeepLynx data warehouse.

Brownlee, JarenM↗

Faraday: A High-temperature Electrolysis Data Explorer

Faraday is a high-temperature electrolysis data visualization tool, which reveals the performance of various button cells under test conditions. These tests and the resulting analytics on their data constitute a state of the industry as the US Department of Energy pushes for the production of hydrogen. Faraday leverages the Idaho National Laboratory's DeepLynx data warehouse to standardize and query button cell data. Faraday programmatically accesses this data in DeepLynx by traversing the schema, represented by a custom ontology. The user interface queries DeepLynx for timeseries data associated with specific button cells in the warehouse, and renders them using JavaScript charts. Additional charting and data analysis techniques are made possible by an auxiliary Python server.

Woodruff, Nathan↗

Digital Twin for Optimizing Real-time Economy of the Integrated Energy Systems

Economic and safe operation of integrated energy systems (IES) requires real-time optimization (RTO) of the control and actions conducted on each system component. In this regard, digital twins (DTs), which consist of a physical system, a virtual system, and the data communication that occurs between the two, are essential for effective RTO. Through the data warehouse, the virtual system is constantly updated with real-time data from the physical system, and functions as the model in the optimization framework. The reduced-order model of the dynamic process model in the virtual system is used in the optimization framework. The optimization results are then returned, via the data warehouse, as control actions to the physical system. This work demonstrates the software capabilities of DT assets for an IES in the context of preparing a DT for an experimental system comprised of Idaho National Laboratory (INL)’s Thermal Energy Delivery System and battery system. For the virtual demonstration, the DTs encompass (1) a physical system, including the Modelica models of the Thermal Energy Delivery System and the battery system; (2) virtual optimization via the Optimization of Real-Time Capacity Allocation (ORCA) platform; and (3) the open-source data warehouse software DeepLynx. This work assesses the performance of ORCA, which utilizes a reduced-order model built using the Risk Analysis Virtual Environment (RAVEN) and trained on the Modelica models and real-time data pipeline through the graph database hosted in DeepLynx. The proposed optimization workflow will be an RTO model based on DTs and the data they generate.

25 ENERGY STORAGE↗

Jester

Jester is a Rust CLI designed to package and send time series or tabular data to the data warehouse DeepLynx. It primarily reads .csv files and then sends those via HTTP or Websocket requests to an external instance of DeepLynx. It is meant to run on a host computer which has access to the data.

Darrington, John↗

An Autonomous Critical Data Extrapolator for the AGN-201m

Nuclear nonproliferation serves as a key goal, being undertaken by the International Atomic Energy Agency (IAEA). To recognize proliferation there are two pathways that states, who intend to use nuclear material for malicious purposes can take, diversion can misuse. Diversion is when fissile nuclear material is declared to the IAEA for non-weapon purposes, but then covertly removed. If the source of nuclear material, that is not declared and not fissionable, is placed inside the reactor core to create fissile material used to create weapons then the state is using the second pathway of proliferation, misuse. With the emerging development in areas of simulation and machine learning the creation of virtual models of reactor systems, digital twins, serve as a potential method to identify proliferation through detecting anomalous behavior in the reactor. A digital twin for a physical nuclear reactor has never been developed, as digital twins serve as an emerging technology. To investigate the process for development and use of a digital twin for a nuclear reactor Idaho State University’s AGN-201m serves as the nuclear reactor used for development of this digital twin. A data acquisition system has been installed to the reactor system allowing for the transfer of collected data from a reactor operation to Idaho National Laboratory’s Deeplynx data warehouse. When utilizing data to train reactor physics and machine learning models, a significant challenge encountered is the initial state of the data. Nuclear proliferation will have the capacity to be detected when the reactor immediately starts up, nor will it occur after the reactor shuts down. Generally, it will be detected when the reactor is operating at some desired power over a sufficient period for that specific reactor design. For the AGN-201m this will be when the reactor is critical (generally 1 mW or above) for a timespan that is within or less than the range of a regular business day. Datasets sent to Deeplynx have had to be manually cut to when the reactor is critical based on plots of power levels. This method is inefficient and laborious, especially when using multiple datasets at once to train a model. To provide a more streamlined approach an automated critical data extrapolator is developed, with capabilities of recognizing when the reactor operation first reaches criticality, and when the reactor undergoes a SCRAM and is shutdown.

99 GENERAL AND MISCELLANEOUS↗

Latency Analysis of the Nexus Digital Twin Framework

Real-time digital catalogs are increasingly relied upon to track metadata and connect disparate data sources for cloud-based data integration efforts. One such tool, Deeplynx Nexus is supporting real-time digital twin efforts through event-driven data integration and time-series queries. Nexus’s usefulness for these applications depends critically on how quickly individual records can be uploaded and downloaded, since delays directly affect the responsiveness of any system built on top of it. However, the actual latency a user should expect from Nexus has not been systematically measured before, particularly for the small, frequent transactions typical of live sensor feeds. Here we show that single-record round-trip latency is 61.1 ms on a local Nexus instance and 391.7 ms on the hosted production infrastructure, a roughly 6.4x difference driven primarily by fixed per-request overhead rather than data volume. This overhead dominates at small scale: comparing single-record and ten-record trials suggests approximately 56 ms of each single-record request is fixed connection and authentication cost rather than data-transfer time, meaning batching even a handful of records is substantially more efficient than transmitting them individually. At large batch sizes, this pattern reverses for uploads, which converge to near parity between local and hosted environments by 25,000-50,000 records, while download latency remains persistently 5.7-6.4x slower on hosted infrastructure even at scale. These results suggest that Nexus deployments intended for real-time digital twin applications should prioritize record batching over single-record transactions, and that download-path optimization on hosted infrastructure offers the largest remaining opportunity to reduce latency at scale. We anticipate these baseline measurements will serve as a reference point for future digital twin projects evaluating whether Nexus’s latency profile meets their real-time requirements, and as a benchmark for tracking the effect of future infrastructure or API changes.

99 - GENERAL AND MISCELLANEOUS↗

Optimization Of Real-time Capacity Allocation

ORCA is a modeling toolset to accelerate real-time control and optimization of digital twins, including virtual models of facilities, physical facilities, and interconnections to allow optimal control of physical facilities using virtual models. ORCA is enabled by INL's RAVEN and DeepLynx software codes.

Talbot, Paul [Idaho National Laboratory] (00000002↗

Agn-201 Digital Twin

This is the repository for all code related to the AGN-201 Nuclear Reactor Digital Twin at Idaho State University. The goal of this code repository is to consolidate all pieces required to run the AGN-201 Digital Twin in the [DeepLynx](https://github.com/idaholab/Deep-Lynx) ecosystem. This is the first successfully launched digital twin of a fissile nuclear reactor that we are aware of. While the code is not complex, the problems of networking, policy, and initial groundwork were significant to overcome.

Darrington, JohnW.↗

Autonomous Anomaly Detection For Continuous Streams

The code implements the Isolation Forest (IFML) algorithm within the digital twin (DT) of the AGN-201 nuclear reactor. The DT captures real-time operational data including control rod positions, reactor power, and temperature. The IFML model isolates anomalies by detecting patterns that deviate from expected operational behavior. The algorithm recursively partitions the data and assigns anomaly scores based on the isolation of rare and different events. By tuning parameters specific to the reactor’s operational data, the IFML identifies deviations such as unauthorized material insertions or reactor reactivity shifts. The system streams data using LabView and integrates with the DeepLynx data warehouse for anomaly processing.

Trevino, Eduardo↗