Engineering PapersSearch

SEARCH · Engineering Papers

Results for “Managing Complexity”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

VA EDH Advanced Software Pipeline Framework Report: Enhancing Automation and Scalability

The VA Environmental Determinants of Health (EDH) Advanced Software Pipeline Framework is designed to enhance the efficiency, scalability, and security of geospatial data processing workflows. This framework integrates modern data orchestration and containerization technologies, including Prefect for workflow automation, Docker for containerization, and PostgreSQL/PostGIS for geospatial data storage and analysis. It ensures standardized, reproducible, and automated data processing, supporting VA objectives related to substance use risk assessment and recovery research. The pipeline addresses key scalability and performance challenges through horizontal and vertical scaling, high-performance computing (HPC) integration, parallel processing, task caching, and dynamic resource allocation. These optimizations improve throughput and reduce latency, allowing the system to efficiently manage large and complex datasets. Additionally, security and compliance measures—such as data encryption (SSL), Role-Based Access Control (RBAC), and adherence to GDPR and HIPAA standards—safeguard sensitive information throughout data transmission and storage. A key implementation of this framework includes the automation of shelter list geolocation workflows, ensuring that up-to-date data is readily available for VA decision-making. Lessons learned from this project include the transition from in-memory processing to incremental storage writes, improving resource management and reliability. Future enhancements aim to expand automation, integrate AI-driven anomaly detection, and incorporate high-performance computing resources. This framework provides a scalable, secure, and adaptable solution for managing geospatial datasets, reinforcing the VA’s ability to support clinical and strategic initiatives through data-driven decision-making.

97 MATHEMATICS AND COMPUTING

Automatic building energy model development and debugging using large language models agentic workflow

Building energy modeling (BEM) is a complex process that demands significant time and expertise, limiting its broader application in building design and operations. While Large Language Models (LLMs) agentic workflow have facilitated complex engineering processes, their application in BEM has not been specifically explored. This paper investigates the feasibility of automating BEM using LLM agentic workflow. Here, we developed a generic LLM-planning-based workflow that takes a building description as input and generates an error-free EnergyPlus building energy model. Our robust workflow includes four core agents: 1) Building Description Pre-Processing, 2) IDF Object Information Extraction, 3) Single IDF Object Generator Suite, and 4) IDF Debugging Agent. These agents divide the complex tasks into manageable sub-steps, enabling LLMs to generate accurate and reliable results at each stage. The case study demonstrates the successful translation of a building description into an error-free EnergyPlus model for the iUnit modular building at the National Renewable Energy Laboratory. The effectiveness of our workflow surpasses: 1) naive prompt engineering, 2) other LLM-based workflows, and 3) manual modeling, in terms of accuracy, reliability, and time efficiency. The paper concludes with a discussion on the interplay between foundational models and LLM agent planning design, advocating for the use of fine-tuned, specialized models to advance this field.

97 MATHEMATICS AND COMPUTING

Constellation: The autonomous control and data acquisition system for dynamic experimental setups

The operation of instruments and detectors in laboratory or beamline environments presents a complex challenge, requiring stable operation of multiple concurrent devices, often controlled by separate hardware and software solutions. These environments frequently undergo modifications, such as the inclusion of different auxiliary devices depending on the experiment or facility, adding further complexity. The successful management of such dynamic configurations demands a flexible and robust system capable of controlling data acquisition, monitoring experimental setups, enabling seamless reconfiguration, and integrating new devices with limited effort. This paper presents Constellation, a flexible and network-distributed control and data acquisition software framework tailored to laboratory and beamline environments, that addresses the limitations of existing solutions. The framework is designed with a focus on extensibility, providing a streamlined interface for instrument integration. It supports efficient system setup via network discovery mechanisms, promotes stability through autonomous operational features, and provides comprehensive documentation and supporting tools for operators and application developers such as controllers and logging interfaces. At the core of the architectural design is the autonomy of the individual components, called satellites, which can make independent decisions about their operation and communicate these decisions to other components. This paper introduces the design principles and framework architecture of Constellation, presents the available graphical user interfaces, shares insights from initial successful deployments, and provides an outlook on future developments and applications.

Autonomy

A Digital Twin Framework Utilizing Machine Learning for Robust Predictive Maintenance: Enhancing Tire Health Monitoring

We introduce a novel digital twin (DT) framework for the predictive maintenance of long-term physical systems. Using monitoring tire health as an application, we show how the DT framework can be used to enhance automotive safety and efficiency, and how the technical challenges can be overcome using a three-step approach. First, to manage the data complexity over a long operation span, we employ data reduction techniques to concisely represent physical tires using historical performance and usage data. Relying on these data, for fast real-time prediction, we train a transformer-based model offline on our concise dataset to predict future tire health over time, represented as remaining casing potential (RCP). Based on our architecture, our model quantifies both epistemic and aleatoric uncertainties, providing reliable confidence intervals around predicted RCP. Second, to incorporate real-time data, we update the predictive model in the DT framework, ensuring its accuracy throughout its lifespan with the aid of hybrid modeling and the use of the discrepancy function. Third, to assist decision-making in predictive maintenance, we implement a tire state decision algorithm, which strategically determines the optimal timing for tire replacement based on RCP forecasted by our transformer model. This approach ensures that our DT accurately predicts system health, continually refines its digital representation, and supports predictive maintenance decisions. Furthermore, our framework effectively embodies a physical system, leveraging big data and machine learning (ML) for predictive maintenance, model updates, and decision-making.

advanced computing infrastructure

Carbon Capture, Transport, And Storage (CTS) Cost Modeling Of The Onshore Gulf Coast

The onshore Gulf of Mexico region presents significant opportunities for CO2 capture, transport, and storage due to its numerous CO2 sources, such as power plants, refineries, and its substantial CO2 storage potential. However, operators face critical decisions in designing an efficient and cost-effective CO2 pipeline network. This study examines the economic implications of two primary strategies: constructing a trunkline with excess initial capacity versus developing dedicated pipelines incrementally as new CO2 sources come online. Building a trunkline first offers the advantage of future-proofing the network, allowing for the accommodation of increased CO2 volumes from various sources over time. However, this approach incurs higher upfront costs and risks underutilizing the transport capacity in the initial stages, potentially resulting in economic inefficiencies. Conversely, constructing dedicated pipelines for each new CO2 source as it becomes operational may avoid the initial overcapacity issue but fails to capitalize on the economies of scale. This could lead to higher overall costs due to the duplication of infrastructure and increased complexity in network management. This research employs a comprehensive cost-benefit analysis, integrating factors such as capital expenditure, operational costs, projected CO2 volumes, and potential economies of scale. Through this analysis, we aim to provide operators with insights into the most economically viable strategy for CO2 pipeline network design in the region. The findings underscore the importance of strategic planning and highlight the trade-offs between immediate capacity utilization and long-term cost savings, ultimately guiding stakeholders towards informed decision-making in the development of CO2 transport infrastructure. Presented at the 41st USAEE/IAEE North American Conference, 3-6 November 2024, Baton Rouge, LA, United States.

Shih, Chung Yan

Energy-Transit Nexus Tools for Bus Fleet Electrification (NEXTBUS)

NEXTBUS is an open-source software project that integrates NLR's bus energy modeling and simulation tools with multi-objective optimization for fleet operations. NLR is collaborating with a transit technology startup, ReVolt, to commercialize these capabilities by deploying NEXTBUS in ReVolt's software platform. The goal is to manage the added complexities of running a heterogeneous fleet, encompassing battery electric and diesel buses, across a large, multi-depot transit network.

33 ADVANCED PROPULSION SYSTEMS

Tracking animal movements via collaborative acoustic telemetry networks: Multiscale habitat use, phenology, and management insights

Abstract Estuaries support diverse fish and invertebrate communities, including resident species that rely on estuarine habitats year‐round and transient migratory species. The unique movement patterns of these animals connect habitats within and far beyond the estuary and are integrally linked to fisheries management objectives. With a focus on Chesapeake Bay, this study leveraged data from collaborative acoustic telemetry networks in the northwest Atlantic to assess habitat use and phenology of movements for seven species of fish (cownose rays, dusky sharks, smooth dogfish, alewife, striped bass, common carp, and blue catfish) and one invertebrate (horseshoe crabs). A total of 288 acoustically tagged individuals were detected >3.2 million times (6,743 to 2,095,717 detections per species) on receivers across ~20.5 degrees of latitude spanning the North American Atlantic seaboard from Florida, USA, to New Brunswick, Canada. Common metrics of movement and phenology grouped these species as resident (common carp, blue catfish, horseshoe crabs), primarily resident in estuaries (juvenile striped bass), and coastal migrant (cownose rays, dusky sharks, smooth dogfish, alewife); maximum distance traveled varied by three orders of magnitude among these species. Further analysis of phenology for coastal migrants elucidated the timing and duration of these species' use of Chesapeake Bay. Collectively, movements linked habitats within Chesapeake Bay and connected the estuary to coastal ecosystems both to the north (e.g., alewife) and south (e.g., cownose rays), creating networks of fisheries management jurisdictions that varied in complexity and identified opportunities for enhancement to current management or co‐management of some species. Our results elucidate the importance of estuaries to species with diverse movement behaviors, identify scales and pathways of habitat connectivity via animal movements, and highlight the utility of collaborative acoustic telemetry networks for quantifying movements relevant to both ecological research and fisheries management.

Livernois, Mariah C.

pdas-experiments

SAND2025-04589O pdas-experiments automates computational experiments of fluid flow simulations. It uses the pressio-demoapps-schwarz package as a basis to break down complex simulations into smaller, manageable parts. This application is an extension of the Sandia Pressio software which uses domain decomposition to work with complex simulations more efficiently. Users can test different simulation setups, while keeping a detailed record of their experiments so they can be reproduced later. The software includes a C++ program that runs individual experiments based on user-defined settings in a YAML file, as well as a Python script that can manage multiple simulations at once. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Tezaur, Irina [Sandia National Lab. (SNL-CA), Live

Water at the Research and Education Complex [Poster]

Purpose: To identify the management methods for stormwater, sanitary sewer, and potable water throughout the facilities at the Research and Education Complex (REC) using drawings and written documents. To confirm the locations and status of wells located throughout the REC. To compile appropriate documentation that relates to water for each facility at the REC, verify that documentation is accurate, and create any new documentation that is necessary through field verification.

54 ENVIRONMENTAL SCIENCES

Enabling HPC Scientific Workflows for Serverless

The convergence of edge computing, big data analytics, and AI with traditional scientific calculations is increasingly being adopted in HPC workflows. Workflow management systems are crucial for managing and orchestrating these complex computational tasks. However, it is difficult to identify patterns within the growing population of HPC workflows. Serverless has emerged as a novel computing paradigm, offering dynamic resource allocation, quick response time, fine-grained resource management and auto-scaling. In this paper, we propose a framework to enable HPC scientific workflows on serverless. Our approach integrates a widely used traditional HPC workflow generator with an HPC serverless workflow management system to create benchmark suites of scientific workflows with diverse characteristics. These workflows can be executed on different serverless platforms. We comprehensively compare executing workflows on traditional local containers and serverless computing platforms. Our results show that serverless can reduce CPU and memory usage respectively by 78.11% and 73.92% without compromising performance.

Andrei da silva, Anderson

Prediction of Distributed River Sediment Respiration Rates Using Community-Generated Data and Machine Learning

River sediment microbial respiration is a key indicator of ecosystem functioning and the biogeochemical fluxes across this critical zone link surface and subsurface waters. As such, there is tremendous interest in measuring and mapping these respiration rates. Respiration observations are expensive and labor intensive; there is limited data available to the community. An open science, collaborative initiative is collecting samples for respiration rate analysis and multi-scale metadata; this evolving data set is being used for making machine learning (ML) predictions at unsampled sites to help inform continued community engagement. However, it is a challenge to find an optimum configuration for ML models to work with this feature-rich (i.e., 100+ possible input variables) data set. Here, we present results from a two-tiered approach to managing the analysis of this complex data set: (a) a stacked ensemble of models that automatically optimizes hyperparameters and manages the training of many models and (b) feature permutation importance to detect the most important features in the models. The major elements of this workflow are modular, portable, open, and cloud-based thus making this implementation a potential template for other applications. The models developed here predict that sediment organic matter chemistry is one of the most important features for predicting sediment respiration rate. Other larger-scale, important features fall into the categories of climatic, ecological, geological, and fluvial settings. Leveraging these larger-scale features to generate data-driven estimates of river sediment respiration rates reveals spatially consistent but heterogeneous patterns across the river network of the Columbia River Basin.

54 ENVIRONMENTAL SCIENCES

POWER ELECTRONICS GRID TIED SYSTEM FINAL REPORT

Across the country, electric utilities are grappling with the persistent hurdles of integrating Distributed Energy Resources (DERs). Managing these assets safely and effectively is a complex endeavor, complicated by varying ownership structures, management philosophies, and the diversity of the technologies themselves. Consequently, the industry has seen a proliferation of bespoke system designs, control strategies, and communication frameworks—forcing utilities to spend significant time and resources developing one-off integration solutions. This project addressed these integration hurdles through a scalable demonstration of intelligent devices designed to coordinate and control diverse resources in low-voltage applications. This concept minimized the need for complex integration by transforming the separate DERs into a dispatchable virtual power plant (VPP) with integrated resiliency functions (called a Node). By collaborating with a utility partner, the project focused on developing rapidly implementable use cases that bridged the gap between theoretical control and real-world deployment

99 GENERAL AND MISCELLANEOUS

OPEN-Augmented Reality GUI for Bioenergy Crop Phenotyping and Precision Agriculture (Donald Danforth Plant Science Center Final Scientific Technical Report)

The project led by the Donald Danforth Plant Science Center, in collaboration with Arizona State University, George Washington University, and Saint Louis University, has made significant strides in advancing the phenotypic analysis of bioenergy crops through the development of an innovative AI processing pipeline. This initiative was primarily funded by ARPA-E, with additional cost-sharing provided by the participating institutions. The project successfully utilized a variety of sensors—3D scanners, thermal, RGB, and hyperspectral—to refine algorithms for data-driven trait signature identification and improve the classification and visualization of plant traits. The developed AI processing pipeline is capable of handling the complex, multidimensional data characteristic of dynamic agricultural environments. 1) Contributions to understanding: The research has advanced the field of plant phenomics by showcasing the synergistic use of various sensor data to enhance the precision of trait analysis in bioenergy crops. Through the integration of 3D scanners, thermal, RGB, and hyperspectral sensors, the project has developed robust data-driven trait signature algorithms and visualization techniques. These innovations have facilitated detailed monitoring and management of plant traits, providing vital insights into plant growth dynamics and stress responses. Further, the project has broadened our understanding of how machine learning can be effectively applied in multi-sensor environments to refine trait analysis. By leveraging diverse datasets, the research has not only improved the accuracy of phenotypic assessments but also established a versatile methodological framework that can be extended beyond agriculture to other fields requiring detailed phenotypic analysis. 2) Technical effectiveness and economic feasibility: The AI processing pipeline developed in this project demonstrated significant technical effectiveness, achieving high throughput analysis of extensive phenotypic data and meeting targeted accuracies. This system exemplified the capability of advanced machine learning technologies to efficiently manage and analyze large, complex datasets. Economically, the implementation of the project-developed pipelines may offer substantial cost savings across multiple sectors. It enhances data analysis processes and significantly reduces the need for manual data interpretation, thereby decreasing both the time and resources required. 3) Public benefit: The project has significantly broadened the scope of agricultural methodologies to enhance phenotypic analysis, with potential applications in various sectors beyond agriculture. Additionally, the initiative fostered an enriching educational and collaborative environment, significantly enhancing the technical skills of participants. It also made substantial contributions to the scientific community by providing open-access data sets and tools, encouraging ongoing research and development across various disciplines. Overall, the project not only met its scientific goals but also showcased the extensive utility of integrating advanced machine learning and sensor data analysis technologies. These advancements have proven instrumental in driving forward both theoretical research and practical applications, setting a strong foundation for future explorations and innovations in data-driven science.

60 APPLIED LIFE SCIENCES

Preparation of the Multi-Site Data Processing at the Vera C. Rubin Observatory

The Vera C. Rubin Observatory’s Legacy Survey of Space and Time (LSST) Camera is scheduled to start taking data in the summer of 2025. The Data Release Production will run the LSST Science Pipe software at data facilities in the US, France and the UK. The LSST Science Pipeline consists of complex directed acyclic graphs (DAGs) of tasks. Rubin will use the Production and Distributed Analysis (PanDA) workflow and workload management system to orchestrate this complex workflow and the distribution of workloads to the data facilities. When run end-to-end by a team of data production staff, this processing (the Science Pipelines, distributed by the workflow and workload management system) is referred to as a 'campaign'. This paper describes the central services and data facility specific services that support this multi-site data process model, including the service deployment infrastructure, the workload and workflow system, the Campaign Management tools, and connection to Rubin Data Management. This paper will also mention the experience of processing the Rubin Commissioning Camera data. All these are part of the effort to scale up the processing capabilities for the expected very large data volume from the LSST Camera.

Yang, Wei [SLAC]

Towards Next-Generation Urban Decision Support Systems through AI-Powered Construction of Scientific Ontology Using Large Language Models—A Case in Optimizing Intermodal Freight Transportation

The incorporation of Artificial Intelligence (AI) models into various optimization systems is on the rise. However, addressing complex urban and environmental management challenges often demands deep expertise in domain science and informatics. This expertise is essential for deriving data and simulation-driven insights that support informed decision-making. In this context, we investigate the potential of leveraging the pre-trained Large Language Models (LLMs) to create knowledge representations for supporting operations research. By adopting ChatGPT-4 API as the reasoning core, we outline an applied workflow that encompasses natural language processing, Methontology-based prompt tuning, and Generative Pre-trained Transformer (GPT), to automate the construction of scenario-based ontologies using existing research articles and technical manuals of urban datasets and simulations. From these ontologies, knowledge graphs can be derived using widely adopted formats and protocols, guiding various tasks towards data-informed decision support. The performance of our methodology is evaluated through a comparative analysis that contrasts our AI-generated ontology with the widely recognized pizza ontology, commonly used in tutorials for popular ontology software. We conclude with a real-world case study on optimizing the complex system of multi-modal freight transportation. Our approach advances urban decision support systems by enhancing data and metadata modeling, improving data integration and simulation coupling, and guiding the development of decision support strategies and essential software components.

96 KNOWLEDGE MANAGEMENT AND PRESERVATION

Dual-Bed Radioiodine Capture from Complex Gas Streams with Zeolites: Regeneration and Reuse of Primary Sorbent Beds for Sustainable Waste Management

Dual-sorbent systems are proposed for radioiodine management with a regenerated primary bed for multiple cycles of use in complex conditions and a secondary bed for disposal with higher waste loadings. Sorbent approaches for the effective capture of gaseous radioiodine (isotopes 129 I and 131 I) produced from a range of nuclear processes have been studied for over half a century. (1−5) Whether or not a sorbent (e.g., molecular sieve) is required to physically screen/trap or chemically bind a radionuclide of interest through chemisorption, the complexity of the gas stream has a large impact on the performance (e.g., loading capacity, selectivity) and active life of a sorbent bed. (3) Silver mordenite (AgZ), the U.S. Department of Energy baseline sorbent for radioiodine capture from nuclear processes, performs well within acidic conditions and at elevated temperatures (6) and can be consolidated into a chemically durable waste form for long-term disposal. (7,8) However, new sorbents are being sought because optimal capture performance of AgZ significantly decreases in dynamic oxidizing environments with competing species, and it is expensive and it contains Ag (a toxic metal). (9) Until a new sorbent is found to replace AgZ, the regeneration and reuse of AgZ is an attractive alternative to a single-use primary sorbent bed. In this regard, a primary sorbent could be designed for enhanced capture in complex gas streams and the ability to be regenerated for reuse. Here, a secondary sorbent could then be tailored for maximum iodine loading in the gas stream and chemical durability within a disposal facility.

chemisorption

Bridging molecular-scale interfacial science with continuum-scale models

Solid–water interfaces are crucial for clean water, conventional and renewable energy, and effective nuclear waste management. However, reflecting the complexity of reactive interfaces in continuum-scale models is a challenge, leading to oversimplified representations that often fail to predict real-world behavior. This is because these models use fixed parameters derived by averaging across a wide physicochemical range observed at the molecular scale. Recent studies have revealed the stochastic nature of molecular-level surface sites that define a variety of reaction mechanisms, rates, and products even across a single surface. To bridge the molecular knowledge and predictive continuum-scale models, we propose to represent surface properties with probability distributions rather than with discrete constant values derived by averaging across a heterogeneous surface. This conceptual shift in continuum-scale modeling requires exponentially rising computational power. By incorporating our molecular-scale understanding of solid–water interfaces into continuum-scale models we can pave the way for next generation critical technologies and novel environmental solutions.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Integrated Distribution Planning

The contemporary distribution planning landscape is comprised of an increasing number of factors that require integration into the engineering of the modern electric grid. Expectations for electric utilities to accommodate heightened awareness of stakeholders' interest in things like decarbonization, resilience and equity are growing. As these interests are formed into objectives, many jurisdictions will experience increasing levels of load modifying technologies like DER, building and industrial electrification and electric vehicles which prove not only to challenge the capabilities of the grid; but the processes by which planning for it is traditionally done. Other related factors that strain the conventional distribution planning mold are the swelling amount and sources of data associated with these technologies and the need it creates for improved capabilities in the processes and tools that manage it. As the complexity of the distribution system expands, so will the distribution system's effects on the transmission and generation systems that it is a part of. Forecasting distribution system load and DER are examples of areas where this complexity will manifest, and harmonizing distribution forecasting with transmission and generation forecasting requires higher amounts of intentionality as these typically separate processes become a solitary one. Of course, core activities do not cease as a utility begins to integrate these other factors, and in this webinar we explore specifics of how distribution planning can be expected to evolve as progress towards Integrated Distribution System Planning is made.

24 POWER TRANSMISSION AND DISTRIBUTION