Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “tasking”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Evaluation of Howard A. Hanson Dam Juvenile Fish Passage and Survival Study: Downstream Passage and Survival (Acoustic Telemetry) (Task Final Report)

The Downstream Passage and Survival (Acoustic Telemetry) task was one of several tasks for the Evaluation of Howard A. Hanson Dam (HAHD) Juvenile Fish Passage and Survival study. For this task, a juvenile fish passage and survival study was conducted at HAHD from March 10, 2024 through June 27, 2024 for arrays in the HAHD Reservoir and Green River, and through July 9, 2024 for the arrays in the Duwamish River. Fish releases occurred over four times from March through May. Active tag (acoustic telemetry) technology was used to provide biological baseline information for the passage and survival of juvenile Puget Sound (PS) Chinook salmon (Oncorhynchus tshawytscha). Results will be used by biologists, engineers, resource managers, and regional decision-makers to inform the engineering design of the new fish passage facility (FPF) at HAHD. However, hatchery sub-yearling PS Chinook salmon of the appropriate size for tag implantation during the study period (> 95 mm in fork length, Geist et al. 2018) were unavailable. Therefore, yearling Coho salmon (O. kisutch) were used as a surrogate species for this acoustic telemetry study.

13 HYDRO ENERGY↗

Framework for Extensible, Asynchronous Task Scheduling (FEATS) in Fortran

Most parallel scientific programs contain compiler directives (pragmas) such as those from OpenMP, explicit calls to runtime library procedures such as those implementing the Message Passing Interface (MPI), or compiler-specific language extensions such as those provided by CUDA. By contrast, the recent Fortran standards empower developers to express parallel algorithms without directly referencing lower-level parallel programming models. Fortran’s parallel features place the language within the Partitioned Global Address Space (PGAS) class of programming models. When writing programs that exploit data-parallelism, application developers often find it straightforward to develop custom parallel algorithms. Problems involving complex, heterogeneous, staged calculations, however, pose much greater challenges. Such applications require careful coordination of tasks in a manner that respects dependencies prescribed by a directed acyclic graph. When rolling one’s own solution proves difficult, extending a customizable framework becomes attractive. The paper presents the design, implementation, and use of the Framework for Extensible Asynchronous Task Scheduling (FEATS), which we believe to be the first task-scheduling tool written in modern Fortran. We describe the benefits and compromises associated with choosing Fortran as the implementation language, and we propose ways in which future Fortran standards can best support the use case in this paper.

Richardson, Brad↗

Task 2.3: Continuous Counter-Current Chromatography

This task in the Separations Consortium focuses on the development of counter-current chromatography (CCC), which is an advanced liquid-liquid chromatography method. Today, CCC is primarily practiced in batch mode operation, but continuous operation will be required for at-scale deployment in a biorefinery setting. In this task, we are developing a continuous CCC process in collaboration with a company that builds CCC units. This work will be demonstrated on the biorefining challenge of lignin valorization. Specifically, we demonstrate use of continuous CCC for both monomer-monomer and monomer-oligomer separations in multiple lignin streams of relevance to the biorefining industry. Breakthroughs in this area would be useful for BETO goals in both sustainable aviation fuel and biochemicals production, including directly contributing to BETO's 2030 lignin valorization goal. Our approach includes developing new computational modeling approaches to optimize both batch and continuous CCC processes, experimental work to determine optimal solvent systems for multiple lignin streams sourced from BETO-funded projects and industrial collaborators, and techno-economic analysis and life cycle assessment to identify the most impactful areas to ultimately enable this approach in the biorefinery. This task overall will enable high-resolution, multi-component, continuous separations at scale, demonstrated on a grand challenge biorefining problem.

BIOMASS FUELS↗

Systems and methods for automatic data management for an asynchronous task-based runtime

A compilation system can define, at compile time, the data blocks to be managed by an Even Driven Task (EDT) based runtime/platform, and can also guide the runtime/platform on when to create and/or destroy the data blocks, so as to improve the performance of the runtime/platform. The compilation system can also guide, at compile time, how different tasks may access the data blocks they need in a manner that can improve performance of the tasks.

Baskaran, Muthu Manikandan↗

Task 5: Developing a Tool to Quantify Liability of Geologic Carbon Storage

This talk provides an overview of work being done on NRAP Phase 3 Task 5. Task 5 involves the liability or cost of responding to potential adverse events. The focus of Task 5 is on the cost of responding to potential leakage of CO2 and brine out of the storage formation or induced seismic incidents.

Morgan, David↗

Classic and Quantum Task-Based Intelligent Runtime for QIRs Running on Multiple QPUs

High-performance computing systems are rapidly evolving into heterogeneous platforms that fuse quantum accelerators with traditional classical processing units (CPUs) and graphical processing units (GPUs). This convergence calls for runtimes capable of managing both classical and quantum workloads in a unified manner. We introduce an intelligent, task-based runtime that marries the Intelligent RuntIme System (IRIS) asynchronous scheduler with a quantum programming stack through the Quantum Intermediate Representation Execution Engine (QIR-EE). Our design allows programs written in the quantum intermediate representation (QIR) to be dispatched concurrently to a variety of back-ends, including multiple quantum simulators and nascent quantum processors, enabling genuine hybrid execution on a single node. To illustrate its practicality, we partition a 4-qubit and 20-qubit circuit into three sub-circuits using quantum circuit cutting via the QCut library. Each sub-circuit is simulated independently by the QIR-EE driver within IRIS, after which a classical post-processing step merges the simulation results to recover the outcome of the original full-circuit computation. This case study demonstrates how finer task granularity can enable the parallel execution and lower the simulation burden per quantum task while preserving overall accuracy, highlighting the feasibility of our hybrid approach.

Miniskar, Narasinga Rao [ORNL] (ORCID:000000018259↗

Energy-efficient cooperative resource allocation and task scheduling for Internet of Things environments

Offloading Internet of Things (IoT) tasks to the cloud for further processing might not always lead to an optimal execution time, particularly in situations such as resource contention, under-provisioning, over-provisioning, and fragmentation. In addition, dynamically optimizing the number of Virtual Machines (VMs) for resource scheduling in order to meet application requirements remains a major research challenge. Further, existing resource scheduling algorithms focus primarily on minimizing operational costs while maximizing resource sharing and utilization. Considering energy utilization as part of the resource allocation and scheduling process as an optimization objective for maintaining load balancing has often been neglected. To address these challenges and more, we propose a cooperative energy-aware resource allocation and scheduling strategy based on a Technique for Order of Preference by Similarity to Ideal Solution (TOPSIS) multi-criteria decision-making method. Here we used the Grid Workloads Archive dataset to evaluate our proposed approach named TOPREAL. Experimental results with respect to the allocation of VM resources when considering processing a large segment of tasks indicate that TOPREAL outperforms existing algorithms in terms of energy savings, with an average improvement of 40.25%, while maintaining an average improvement of 16.21% when it comes to execution time. Results also demonstrate that our method can save an average of 78.06 processing hours and 63,215kJ of energy when compared to existing scheduling algorithms. These results demonstrate the effectiveness of our proposed model and the viability of using multi-criteria decision-making techniques such as TOPSIS to solve the resource allocation and scheduling problem in edge environments.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Investigating the impact of a multi-module operation environment on the task performance time of human operators – An explanatory study

The worldwide demand for Small Modular Reactors (SMRs) has surged in recent years due to their enhanced safety and versatility in supporting diverse industrial sectors. A unique feature of SMR operation is that a single human operator is responsible for managing multiple modules. Therefore, securing a sufficient amount of human performance data pertaining to this new environment is essential for the safe operation of SMRs. In this explanatory study, a series of experiments were conducted using the NuScale simulator, a representative SMR design, with student operators. A total of 12 student operators were assigned two types of off-normal events and asked to cope with them using paper-based procedures. Subsequently, their task performance times were compared with those of student operators responsible for a single unit based on the Task Complexity (TACOM) measure. Results indicate that the performance of student operators under the experimental conditions of this study degraded by a factor of 2 to 3, depending on the characteristics of the off-normal events.

21 SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLANTS↗

Energy and Emission Prediction for Mixed-Vehicle Transit Fleets Using Multi-task and Inductive Transfer Learning

Public transit agencies are focused on making their fixed-line bus systems more energy efficient by introducing electric (EV) and hybrid (HV) vehicles to their fleets. However, because of the high upfront cost of these vehicles, most agencies are tasked with managing a mixed-fleet of internal combustion vehicles (ICEVs), EVs, and HVs. In managing mixed-fleets, agencies require accurate predictions of energy use for optimizing the assignment of vehicles to transit routes, scheduling charging, and ensuring that emission standards are met. The current state-of-the-art is to develop separate neural network models to predict energy consumption for each vehicle class. Although different vehicle classes’ energy consumption depends on a varied set of covariates, we hypothesize that there are broader generalizable patterns that govern energy consumption and emissions. In this paper, we seek to extract these patterns to aid learning to address two problems faced by transit agencies. First, in the case of a transit agency which operates many ICEVs, HVs, and EVs, we use multi-task learning (MTL) to improve accuracy of forecasting energy consumption. Second, in the case where there is a significant variation in vehicles in each category, we use inductive transfer learning (ITL) to improve predictive accuracy for vehicle class models with insufficient data. As this work is to be deployed by our partner agency, we also provide an online pipeline for joining the various sensor streams for fixed-line transit energy prediction. Here, we find that our approach outperforms vehicle-specific baselines in both the MTL and ITL settings.

97 MATHEMATICS AND COMPUTING↗

Integrating and Characterizing HPC Task Runtime Systems for hybrid AI-HPC workloads

Scientific workflows increasingly involve both HPC and machine-learning tasks, combining MPI-based simulations, training, and inference in a single execution. Launchers such as Slurm’s srun constrain concurrency and throughput, making them unsuitable for dynamic and heterogeneous workloads. We present a performance study of RADICAL-Pilot (RP) integrated with Flux and Dragon, two complementary runtime systems that enable hierarchical resource management and high-throughput function execution. Using synthetic and production-scale workloads on Frontier, we characterize the task execution properties of RP across runtime configurations. RP+Flux sustains up to 930 tasks/s, and RP+Flux+Dragon exceeds 1,500 tasks/s with over 99.6% utilization. In contrast, srun peaks at 152 tasks/s and degrades with scale, with utilization below 50%. For IMPECCABLE.v2 drug discovery campaign, RP+Flux reduces makespan by 30–60% relative to srun/Slurm and increases throughput more than four times on up to 1,024. These results demonstrate hybrid runtime integration in RP as a scalable approach for hybrid AI-HPC workloads.

HPC-AI↗

The IEA Wind Task 49 Reference Floating Wind Array Design Basis

International Energy Agency Wind Technology Collaboration Programme (IEA Wind) Task 49 on Integrated Design of Floating Wind Arrays is an international collaboration aiming to advance the development of large-scale floating wind farms by providing open-access resources to the research and development and planning communities. The work of Task 49 focuses on array-level challenges related to the colocation of many floating wind turbines; their layouts, mooring systems, and cabling systems; failure risks; logistical considerations; marine spatial planning needs; and future research needs and innovation directions. This report provides a general design basis for the development of reference floating wind farm designs. These reference array designs will extend the scope of existing reference floating wind turbine designs to facilitate research on array-level floating wind technology challenges and innovations. The design basis promotes coordination and consistency in developing the reference array designs.

17 WIND ENERGY↗

IEA Wind TCP Task 49: Reference Site Conditions for Floating Wind Arrays

This report, prepared within Work package 1 of IEA Wind Task 49, presents reference site conditions for floating wind arrays to serve as a design basis for the techno-economic design of reference floating wind arrays. Data of the reference sites presented here are publicly available in an open database and thus support fast development and comparable design of floating wind arrays for various relevant conditions. The development of these reference sites drew on existing open access datasets and ongoing research projects of task participants. Six classes were identified that describe relevant key conditions for the design and development of floating wind arrays: met-ocean conditions, seabed conditions, coastal infrastructure, environmental impact, socio-economic impact, as well as regulations and permissions. The reference sites for the techno-economic design of floating wind arrays are based on a concept with building blocks to synthesize purpose-built site representations. In each of the identified classes with influencing design factors, building blocks are used to describe the characteristic properties and their spread. However, the latter three classes (i.e., environmental impact, socio-economic impact, regulations and permissions) are not included in the reference site conditions due to limited knowledge and lack of reliable criteria to quantify their impact on the techno-economic design in numeric parameters. Building blocks with key parameters for the techno-economic design of floating wind arrays are provided for met-ocean conditions, seabed conditions, and coastal infrastructure. For met-ocean conditions, multiple sites were selected for detailed analysis that represent a range of conditions across the pipeline of floating wind projects. Wind conditions and sea states are separated, and each location considers both the severity of wind and waves e.g. one site may have a moderate wave condition but severe wind condition. From this pipeline, eleven representative sites were selected where both site-specific analysis was available within the consortium, and where they represent different parts of the global pipeline. The eleven sites are: Hannibal (Italy), Humboldt (US), Ulsan (South Korea), MoneyPoint One (Ireland), Havbredey (UK), Fukushima (Japan), Utsira Nord (Norway), Gulf of Maine (US), Sud de la Bretagne II (France), Sorlige Nordsjo II (Norway). Each of these sites is summarized in the main report while more details about the studies and analyses behind the datasets are provided in the appendix. For seabed conditions, general information about the geotechnical parameters is provided and a baseline is established for the geotechnical parameters and stratigraphy that may be encountered on the sites. A set of six 'synthetic cases' is defined as building blocks providing the different parameters required for design under each case/soil condition. For the coastal infrastructure, general information about the main requirements is provided that a port should comply with to provide a satisfactory service during the construction of floating offshore wind arrays. Minimum port infrastructural requirements are provided for three types of ports.

17 WIND ENERGY↗

MTL_TX: A Multi-Task Transformer Model for Improved Radiation Time-Series Estimation

Controlling radiation doses at potential radioactive facilities is critical to ensuring the safety of both personnel and the public. At the Thomas Jefferson National Accelerator Facility (JLab), multiple sensors are deployed around the three experimental halls to monitor key parameters, including single-beam current, energy levels, current leakage, and radiation values during accelerator operations. In this study, we developed a Multi-task Transformer model, MTL_TX, to accurately estimate radiation doses at sensor locations based on historical data, with the aim of enhancing safety in accelerator facilities and surrounding public areas. To improve estimation accuracy, we integrated two innovative components into the proposed model: hierarchical feature embedding (HFE) and multi-level decomposition attention (MDA). Additionally, the multi-task learning (MTL) framework effectively leverages correlations among multiple sensors, enabling individual estimations for each sensor. MTL_TX achieved outstanding results on data collected in 2018, with an MSE of 0.1464, an RMSE of 0.2353, and an R 2 score of 0.8584. Furthermore, when trained on 2018 data, MTL_TX exhibited excellent generalization capability to unseen datasets from 2016 to 2019, achieving an MSE of 0.1407, an RMSE of 0.2263, and an R 2 score of 0.8831. These results demonstrate a significant improvement over existing state-of-the-art models.

Transformer↗

Multi-task Parallelism for Robust Pre-training of Graph Foundation Models on Multi-source, Multi-fidelity Atomistic Modeling Data

Graph foundation models using graph neural networks promise sustainable, efficient atomistic modeling. To tackle challenges of processing multi-source, multi-fidelity data during pre-training, recent studies employ multi-task learning, in which shared message passing layers initially process input atomistic structures regardless of source, then route them to multiple decoding heads that predict data-specific outputs. This approach stabilizes pre-training and enhances a model’s transferability to unexplored chemical regions. Preliminary results on approximately four million structures are encouraging, yet questions remain about generalizability to larger, more diverse datasets and scalability on supercomputers. We propose a multi-task parallelism method that distributes each head across computing resources with GPU acceleration. Implemented in the open-source HydraGNN architecture, our method was trained on over 24 million structures from five datasets and tested on the Perlmutter, Aurora, and Frontier supercomputers, demonstrating efficient scaling on all three highly heterogeneous super-computing architectures.

Lupo Pasini, Massimiliano [ORNL] (ORCID:0000000249↗

Asynchronous distributed-memory task-parallel algorithm for compressible flows on unstructured 3D Eulerian grids

Here, we discuss the implementation of a finite element method, used to numerically solve the Euler equations of compressible flows, using an asynchronous runtime system (RTS). The algorithm is implemented for distributed-memory machines, using stationary unstructured 3D meshes, combining data-, and task-parallelism on top of the Charm++ RTS. Charm++’s execution model is asynchronous by default, allowing arbitrary overlap of computation and communication. Task-parallelism allows scheduling parts of an algorithm independently of, or dependent on, each other. Built-in automatic load balancing enables continuous redistribution of computational load by migration of work units based on real-time CPU load measurement. The RTS also features automatic checkpointing, fault tolerance, resilience against hardware failure, and supports power-, and energy-aware computation. We demonstrate scalability up to 25 x 10 9 cells at $\mathscr{O}$10 4 compute cores and the benefits of automatic load balancing for irregular workloads. The full source code with documentation is available at https://quinoacomputing.org.

42 ENGINEERING↗

Micromovements and discomfort associated with flight mission with helmet operation tasks with different levels of cognitive workload

When performing stationary tasks under elevated cognitive workload, individuals must perform continual muscle contractions to maintain stability of the body, resulting in fatigue of the postural muscles. When the muscles perform these contractions in a prolonged manner, the body potentially responds through small changes in body movements—micromovements that may lead to discomfort. The study purpose was to evaluate impact of cognitive load on micromovements. The micromovements were measured during three different cognitive workloads; low, medium, and high. The NASA-TLX score was used to evaluate the perceived mental workload and discomfort was assessed by visual analog scale. In total, 60 subjects (30 males and 30 females) were recruited and performed cognitive tasks that simulated flight operations such as changing the radio frequency based on air traffic control messages, balancing the fuel levels in simulated fuel tanks, and aiming a reticle in a designated moving target using the cyclic control. Cognitive load was defined by the frequency of events. Micromovements were defined by changes in the center of pressure (COP) of the seat pan and COP standard deviation. It was found that the high cognitive workloads had the highest NASA-TLX scores including mental demands, temporal demands, and effort. The neck area had the highest overall levels of discomfort followed by upper back. The highest standard deviation for COP shift and number of micromovements occurred for medium cognitive workloads. In conclusion, while there were some interesting trends, few trends reached a statistical significance due to high variability among subjects for the outcome variables.

60 APPLIED LIFE SCIENCES↗

Online task-space motion control for positioner-coordinated multi-robot manufacturing systems

Incorporating multiple robotic manipulators into large-scale manufacturing systems enhances production efficiency and expands manufacturing capabilities beyond those of single-robot systems. Workpiece positioners in robotic manufacturing have demonstrated significant benefits for process optimization, but coordination strategies for multi-robot systems with shared positioners have received limited attention. This work presents a task-space coordinated trajectory-tracking control framework for multi-robot manufacturing systems, in which robots coordinate their motions within a shared, dynamic workpiece positioning frame. A workpiece positioner actively adjusts the pose of the manufactured component to enable greater operational concurrency and improve overall production efficiency. The proposed motion-coordination scheme employs a distributed and scalable architecture, supporting coordination across heterogeneous multi-robot systems. Two optimization methodologies are introduced to manage kinematic redundancies and maintain continuous, near-optimal operation throughout the manufacturing process. The first strategy exploits a task-space dimensionality reduction to achieve locally optimal configurations by leveraging symmetry-axis rotations of the tool. The second strategy utilizes the workpiece positioner to drive the coordinated robots toward stable and kinematically favorable configurations. For both optimization strategies, multiple objectives are defined to improve key performance metrics, including manipulability, configuration consistency, proximity to mechanical limits, and motion efficiency. Addressing a key limitation of existing coordination approaches, the framework is designed around online setpoint modification, allowing coordinated robots to respond effectively to in-situ process feedback. The proposed control framework is validated using the Robot Operating System (ROS) middleware on a combination of physical and simulated multi-robot system hardware.

Arbogast, Alex [ORNL] (ORCID:0000000154740723)↗

Task-oriented machine learning surrogates for tipping points of agent-based models

We present a machine learning framework bridging manifold learning, neural networks, Gaussian processes, and Equation-Free multiscale approach, for the construction of different types of effective reduced order models from detailed agent-based simulators and the systematic multiscale numerical analysis of their emergent dynamics. The specific tasks of interest here include the detection of tipping points, and the uncertainty quantification of rare events near them. Our illustrative examples are an event-driven, stochastic financial market model describing the mimetic behavior of traders, and a compartmental stochastic epidemic model on an Erdös-Rényi network. We contrast the pros and cons of the different types of surrogate models and the effort involved in learning them. Importantly, the proposed framework reveals that, around the tipping points, the emergent dynamics of both benchmark examples can be effectively described by a one-dimensional stochastic differential equation, thus revealing the intrinsic dimensionality of the normal form of the specific type of the tipping point. This allows a significant reduction in the computational cost of the tasks of interest.

97 MATHEMATICS AND COMPUTING↗