Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “tasking”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

I.3.4.1.1 Overview of Advanced Characterization Within the Powertrain Materials Program (Task 4A1) (Oak Ridge National Laboratory); (Task 4A2) (Argonne National Laboratory); and (Task 4A3) (Pacific Northwest National Laboratory)

This report describes the activities performed during the third year of Thrust 4A, “Advanced Characterization,” within the DOE-EERE VTO PMCP. The goal of the PMCP, which was launched in October 2018, has been to accelerate design, development, demonstration, and deployment of new, cost-effective advanced alloy solutions via a modern ICME approach. The properties of these new materials are targeted to enable improvements in engine efficiency, lightweighting, and durability enhancement over the full range of on-road vehicle classes (e.g., Classes 1-8), including range extenders for future electric HD freight vehicles.

33 ADVANCED PROPULSION SYSTEMS↗

OpenMP Target Task: Tasking and Target Offloading on Heterogeneous Systems

This work evaluated the use of OpenMP tasking with target GPU offloading as a potential solution for programming productivity and performance on heterogeneous systems. Also, it is proposed a new OpenMP specification to make the implementation of heterogeneous codes simpler by using OpenMP target task, which integrates both OpenMP tasking and target GPU offloading in a single OpenMP pragma. As a test case, the authors used one of the most popular and widely used Basic Linear Algebra Subprogram Level-3 routines: triangular solver (TRSM). To benefit from the heterogeneity of the current high-performance computing systems, the authors propose a different parallelization of the algorithm by using a nonuniform decomposition of the problem. This work used target GPU offloading inside OpenMP tasks to address the heterogeneity found in the hardware. This new approach can outperform the state-of-the-art algorithms, which use a uniform decomposition of the data, on both the CPU-only and hybrid CPU-GPU systems, reaching speedups of up to one order of magnitude. The performance that this approach achieves is faster than the IBM ESSL math library on CPU and competitive relative to a highly optimized heterogeneous CUDA version. One node of Oak Ridge National Laboratory’s supercomputer, Summit, was used for performance analysis.

Valero Lara, Pedro↗

On the feasibility of future colliders: report of the Snowmass'21 Implementation Task Force

Colliders are essential research tools for particle physics. Numerous future collider proposal were discussed in the course of the US high energy physics community strategic planning exercise Snowmass'21. The Implementation Task Force (ITF) has been established to evaluate the proposed future accelerator projects for performance, technology readiness, schedule, cost, and environmental impact. Corresponding metrics has been developed for uniform comparison of the proposals ranging from Higgs/EW factories to multi-TeV lepton, hadron and ep collider facilities, based on traditional and advanced acceleration technologies. Here, this article describes the metrics and approaches, and presents evaluations of future colliders performed by the ITF.

43 PARTICLE ACCELERATORS↗

JobQueue-PG: A Task Queue for Coordinating Varied Tasks Across Multiple HPC Resources and HPC Jobs

The software allows for queueing and dispatch of tasks of small, varied, or uncertain runtimes across multiple HPC jobs, resources, and other computing systems. The software was designed to allow scientists to enqueue, run, and accumulate results from computational experiments in an efficient, manageable manner. For example, the software can be used to enqueue many small computational experiments and run them using several long-running multi-node HPC jobs that may or may not run simultaneously.

Tripp, Charles↗

Task 12 PV Sustainability - Status of PV Module Recycling in Selected IEA PVPS Task 12 Countries

Photovoltaic (PV) deployment has accelerated in recent years compared to projections in the early 2010s. This means that PV end of life (EOL) waste streams will also increase at a higher pace than anticipated. To meet and optimise PV EOL management, appropriate regulatory and technological approaches must be implemented in the near term, ensuring that available options are adapted to the conditions of each country or region. This report aims to review the current regulatory and industrial landscape for selected countries belonging to the International Energy Agency's PV Power Systems technology collaboration programme, to assess status of PV EOL management, allow for comparison and cross-fertilization, and establish a foundation for future tracking of progress. Although volumes of EOL PV modules are still small, EOL PV is treated and recycled in a proper manner in the countries and regions that have EOL regulations in place. However, the current low volumes, limited available recycling technologies, logistics challenges, and undeveloped markets for recovered materials result in a high-cost, low-revenue scenario for PV module recycling globally. Nevertheless, the implementation of PV EOL regulations in more countries and R&D investment in PV recycling is expected to accelerate further improvements to meet future demand and to achieve high-value, low-cost recycling. We hope this report contributes to understanding the global status of PV recycling and to accelerating its development as a promising option for the proper EOL management of PV modules in the coming decades.

14 SOLAR ENERGY↗

Overcoming sparse datasets with multi-task learning as applied to high entropy alloys

Abstract The design of novel High Entropy Alloys for use in high-temperature applications is an area of active interest due to their potential to provide exceptional properties compared to conventional alloys. Since the increased popularity of machine learning, an important cog in the design process has been training surrogate models on alloy properties. However, these Single-Task models are trained on individual mechanical properties and do not take advantage of the relatedness between properties. Multi-Task models can capture the interdependencies between tasks, leading to potentially more accurate predictions for all tasks. In this paper, we investigate if Multi-Task models can show improvement over Single-Task models when used for predicting the mechanical properties of these alloys. To ensure fair evaluation between the models, we apply L 0 regularization and skip connections to the models, which allows them to adjust the number of model parameters and depth for optimal performance. We find that the Multi-Task models can leverage task relationships to perform better than Single-Task models, especially for high amounts of missing data in the tasks. Furthermore, adding simple auxiliary targets can boost Multi-Task performance even further despite not being effective as input descriptors to single-task models themselves. We anticipate that the proposed strategies can achieve more accurate predictions and consequently enable better design capabilities for such data-constrained domains without incurring much additional computational cost.

Debnath, Arindam (ORCID:0000000194274499)↗

DECOVALEX-2023: Task F1 Final Report

DECOVALEX-2023 Task F is a comparison of models and methods for post-closure performance assessment (PA) of a deep geologic repository for radioactive waste. The general aims of Task F are to build confidence in the models, methods, and software used for PA and to stimulate additional research and development in PA methodologies. The task objectives are to motivate development of PA modelling skills and capabilities, to examine the influence of model choices on calculated repository performance, and to compare the uncertainties introduced by model choices to other sources of uncertainty. Task F involves no actual experiment or site. It is a PA modelling exercise that requires the conceptual development of hypothetical repository designs and geologic settings. Because three of the teams were interested in salt and the rest of the teams were interested in crystalline rock, Task F was split into two branches: Task F1 for crystalline rock and Task F2 for salt. This report is for Task F1, crystalline rock. Teams from seven countries (Canada, Czech Republic, Germany, Korea, Sweden, Taiwan, and United States) participated in Task F1. The teams worked together to define the features, events, and processes of the reference case repository and established a set of performance measures. In addition, they defined a set of benchmark problems designed to test and compare modelling capabilities for fracture flow and transport at different scales. The repository design and benchmark problems are documented in a Task Specification that evolved over time as the group honed the specifications. The benchmark problems verified that each team can aptly model flow and transport in fractured media in 1-, 2-, and 3-dimensions. Two general approaches were used for the 3-dimensional benchmarks: discrete fracture network (DFN) and equivalent continuous porous medium (ECPM). DFN modelling involves explicit meshing of each fracture while ECPM modelling aims to capture the effective porosity and directional permeability of each cell in a space-filling mesh as affected by intersecting fractures. In some models, a combination of the two is used, i.e., DFN for large known fractures and ECPM for the rest of the domain. Transport is solved by using either the advection-dispersion equation or particle tracking. Although some variation is observed among model breakthrough curves in the benchmark problems, there is strong agreement in breakthrough behaviour up to at least the 75 th percentile for all benchmarks. At the 90 th percentile, breakthrough results show larger differences, suggesting several models retain substantially higher fractions of tracer in regions of slower moving water. In addition to the flow and transport benchmarks, several teams completed the source term benchmark, verifying capabilities for modelling radionuclide decay and ingrowth, waste package breach, instant release fractions, fuel matrix degradation rates, and radionuclide solubility limitations. The reference case is conceptualized as a generic spent fuel repository at a depth of 450 m in fractured crystalline rock. The repository has 50 parallel backfilled drifts, each with 50 deposition holes 6 m apart. Each deposition hole contains a 4-PWR waste package and bentonite buffer. The rock domain is 5 km in length, 2 km in width, and 1 km in depth. It has 6 deterministic fractured deformation zones and a multitude of stochastic fractures. Teams generally used the ECPM approach for the entire rock or a hybrid approach in which the deterministic fracture zones are modelled with a DFN and the rest of the rock is modelled by ECPM. Of the reference case problems specified, only the results of the initial reference case problem are compared in this report. The initial problem focuses on transport from the deposition holes to the surface, i.e., it neglects waste package performance. Tracers are released at all waste package locations at time zero and tracked for their releases to the near field and ground surface. The water fluxes calculated at the ground surface entry and exit regions of the domain are similar for all models except for two that have considerably lower fluxes. For tracer transport, large differences are observed among models in the magnitude of tracer transported. Much of the difference appears to be due to how the repository is implemented and hence the different degrees of repository simplification. Models that exclude the drifts, buffer, and backfill from the domain tend to show greater release of tracers and radionuclides from the repository. The initial study presented here indicates that major differences in modelling important processes within the repository (e.g., diffusion through buffer and backfill) can produce broadly different release and transport results, especially when those processes are excluded. Even for the models that included all specified features, events, and processes, the results show significant differences and demonstrate the importance of examining multiple modelling approaches in performance assessment. The differences in results observed in this study are expected to motivate teams to either increase complexity in future versions of the reference case models or to improve methods to account for the effects of simplified features and processes. Either way, future improvements in these models are expected to produce results that more closely agree.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

Iterative multi-task learning and inference from seismic images

Seismic interpretation aims to extract quantitative and interpretable attributes from a seismic image produced using some migration method to inform characteristics of a subsurface reservoir or target of interest. Current paradigms for computing seismic attributes mostly rely on single-task algorithms. We develop an iterative, multi-task machine learning method to learn and infer multiple attributes from a seismic image. This method is composed of two stages: a multi-task inference stage and a multi-modal, multi-task refinement stage. The basic mechanism of this method is that we train a multi-task inference neural network (NN) to estimate a set of attributes, including a relative geological time (RGT), a denoised higher-resolution (DHR) seismic image, and multiple fault attributes (including probability, dip, and strike), from a low-resolution, noisy seismic image; then we input the inferred attributes to a multi-task refinement NN to enhance the raw inference results iteratively. The two multi-task NNs are trained separately based on synthetic seismic images and associated attributes generated by a geological modeling algorithm. The software we intend to release is a PyTorch implementation of this multi-task learning method for both 2D and 3D cases along with scripts to run the training/validation. The algorithm and software can be a useful tool for automatic seismic interpretation.

Gao, Kai↗

Occupational Experience Effects on Physiological and Perceptual Responses of Common Soldiering Tasks

Abstract Cohen BS, Redmond JE, Haven CC, Foulis SA, Canino MC, Frykman PN, Sharp MA. Occupational Experience Effects on Physiological and Perceptual Responses of Common Soldiering Tasks. J Strength Cond Res 37(4): 894–901, 2023—This study measured the impact of occupational experience (i.e., time spent deployed, in military service, and in job and task performance frequency in training, deployment, and study practice) on the physiological (heart rate [HR] and oxygen consumption [VO 2 ]) and perceptual (rate of perceived exertion [RPE]) responses to performance of critical physically demanding tasks (CPDTs). Five CPDTs (road march, build a fighting position, move under fire, evacuate a casualty, and drag a casualty to safety), common to all soldiers, were performed by 237 active duty soldiers. Linear regression models examined the association between measures of experience and physiological and perceptual performance responses to task demands. The level of significance was adjusted for multiple comparisons and set at ρ ≤ 0.0125 for this study. Significant and notable effect sizes included the impact of time spent deployed on the physiological measures of the road march (PostHR F = 24.84, p < 0.0001, β=-9.65), sandbag fill (PostHR F = 8.26, p = 0.005, β = −2.83), and sandbag carry (MeanHR F = 7.51, p = 0.007, β = −1.12; PostHR F = 7.35, p = 0.007, β = −0.87). For the road march task, there was a nearly 10 bpm decrease in postperformance HR for every year spent deployed. Road march, sandbag fill, and sandbag carry tasks PostHRs were also notably negatively associated with the experience measures of time in their MOS (job and time in military service but not for other physiological and perceptual responses, including VO 2 and RPE. Frequency of task performance in training, deployment, and study practice was not meaningfully associated with experience. The results suggest that increasing task familiarization through on-the-job occupational operational experience may result in greater proficiency and reduced physiological effort.

Sport Sciences↗

DECOVALEX-2023: Task F2 Salt Final Report

The subject of Task F of DECOVALEX-2023 concerns performance assessment modelling of radioactive waste disposal in deep mined repositories. The primary objectives of Task F are to build confidence in the models, methods, and software used for performance assessment (PA) of deep geologic nuclear waste repositories, and/or to bring to the fore additional research and development needed to improve PA methodologies. In Task F2- (salt), these objectives have been accomplished through staged development and comparison of the models and methods used by participating teams in their PA frameworks. Coupled-process submodels and deterministic simulations of the entire PA model for a reference scenario for waste disposal in domal salt have been conducted. The task specification has been updated continuously since the initiation of the project to reflect the staged development of the conceptual repository model and performance metrics. Thermal, hydrological, mechanical, and chemical properties of individual components of the engineered and natural system were chosen for relevance by participating teams. The salt reference case system was characterized using data and measurements collected at relevant underground research laboratories (URLs), field sites, and simulation results from teams with specialized modelling capability. Participating teams made a wide range of model assumptions from compartmentalized networks to full 3D models of the salt formation. No single contributed model includes full-fidelity representation of all the features, events, and processes (FEPs) detailed in the task specification, but almost all features and processes are represented in at least one model. Despite differences in the modelling strategies developed by participating teams, all models indicate that salt compaction and radionuclide diffusion are key processes in the repository, and for the FEPs and model scenario considered, little of the disposed radionuclides will migrate beyond the repository seal over the 100,000 year simulations. In general, the model output quantities have the largest differences over the short term and near the waste. The models tend to be more similar further from waste and at later time. Disparities between the models are believed to be due to differing simplifications from the task specification, some of which are chosen simplifications to reduce complexity, and some are restrictions imposed by the modelling tools. A second round of this task has been accepted for DECOVALEX-2027 in conjunction with Task F1 on crystalline PA modelling. The future round includes waste package heating, improved modelling of salt creep closure, additional comparisons of coupled-process sub-models, and the impact of repository engineering design on radionuclide migration in the repository. Participants will also propose and finalize a set of uncertain inputs for the reference case simulations, propagate these uncertainties in a set of realizations, and conduct sensitivity analyses on the simulation results.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

Molecular To Mesoscale Targeting of Oxoanions with Multi-Tasking Hosts

Achieving a better understanding of anion interactions both in solution and crystalline state was the overarching goal of this project. Anions are everywhere throughout Nature and play important roles in biological and environmental processes. They can be beneficial or deleterious or both in different situations and concentrations. For either reason it is important to have molecules that can bind anions for key needs that benefit society. However, recognition of specific anions is challenging due to the diffuse nature of their negative charge(s) as well as their various shapes and sizes. Understanding the basic properties of anions and how they interact with other molecules and ions in surrounding environments is key to selective recognition. In this project multi-tasking molecules for selective binding of targeted anions were designed to achieve cooperativity and synergism in one rather than multiple host molecules, including (1) cation:anion pair hosts for anions with charges of -2 or greater; (2) pH and redox activated hosts for on-off binding and release; and (3) multiple anion capture in extended host networks. Our design strategy was to combine the use of simple inexpensive building blocks and high yield synthetic pathways to provide economically feasible scale-up for applications. Oxoanions representing multiple shapes and charges were chosen based on having the potential for significant impact on DOE separations needs. Amide/amine-based macrocycles and urea/amine-based chelates and macrocycles with multiple hydrogen bonding sites provided the basic anion-binding frameworks. Successful multi-tasking outcomes were forthcoming in all three tasks. In Task 1, successful ion pair binding for anions with multiple charges was achieved. Furthermore, the ion pair molecules were capable of extended interactions through supramolecular intertwining, like fishing nets for capturing pools of fish (also fitting with Task 3). In Task 2, molecules were synthesized possessing on-off switches. These included a pH sensitive sensor for on-off binding of anions in general, as well as an electrochemical sensor selective for sulfate capture. Three new classes of extended anion host networks capable of binding multiple ions was a major outcome of Task 3. These systems included: anion sensitive, fluorescent organogels; channel-forming macrocycles for studying anion-water including larger macrocyclic cluster sandwiches; and, the offshoot of Task 1, fishing net ion-pair networks for higher valent anions. These strategies can be expanded in the future to other ions and molecules for a better understanding of intermolecular and interionic interactions.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

A Typology of Decision-Making Tasks for Visualization

Despite decision-making being a vital goal of data visualization, little work has been done to differentiate decision-making tasks within the field. While visualization task taxonomies and typologies exist, they often focus on more granular analytical tasks that are too low-level to describe large complex decisions, which can make it difficult to reason about and design decision-support tools. In this paper, we contribute a typology of decision-making tasks that were iteratively refined from a list of design goals distilled from a literature review. Our typology is concise and consists of only three tasks: CHOOSE, ACTIVATE, and CREATE. Although decision types originating in other disciplines exist, we provide definitions for these tasks that are suitable for the visualization community. Our proposed typology offers two benefits. First, the ability to compose and hierarchically organize the tasks enables flexible and clear descriptions of decisions with varying levels of complexities. Second, the typology encourages productive discourse between visualization designers and domain experts by abstracting the intricacies of data, thereby promoting clarity and rigorous analysis of decision-making processes. We demonstrate the benefits of our typology through four case studies, and present an evaluation of the typology from semi-structured interviews with experienced members of the visualization community who have contributed to developing or publishing decision support systems for domain experts. Our interviewees used our typology to delineate the decision-making processes supported by their systems, demonstrating its descriptive capacity and effectiveness. Finally, we present preliminary findings on the usefulness of our typology for visualization design.

97 MATHEMATICS AND COMPUTING↗

DECOVALEX-2023: Task E Final Report

This is the Task E final report for DECOVALEX-2023. Task E is focused on understanding thermal, two-phase hydrological, and mechanical (TH2M) processes, especially related to predicting brine migration in the excavation damaged zone around a heated excavation in salt. Salt is attractive as a disposal medium for radioactive waste because it is self-healing and is essentially impermeable and essentially non-porous in the far field (away from excavations). Investigation of the short-term (days to years) near-field (centimeters to tens of meters) behavior of salt is important for radioactive waste disposal because this early period strongly controls the amount of brine in a salt repository. Brine leads to corrosion of waste forms and waste packages, and possible dissolution of radionuclides with brine transport being a potential transport vector to the accessible environment. The main test case used in Task E is the ongoing Brine Availability Test in Salt (BATS) heater test located underground at the Waste Isolation Pilot Plant (WIPP) near Carlsbad, New Mexico, USA. The Task was divided into a series of Steps. Step 0 was an introduction to processes in salt, that included matching historical unheated brine inflow data from boreholes at WIPP and matching temperature observations during BATS heater test 1a. Step 1 included validation of models against a thermo-poroelastic analytical solution relevant to heated boreholes in salt, and two-phase flow around an excavation in salt. Step 2 required all the individual components covered in steps 0 and 1 to come together to match observed brine inflow behavior during the BATS 1a heater test. There were a range of approaches from the teams, from mechanistic to prescriptive. Given the uncertainties in the problem, some teams used one- or two-dimensional models of the processes, while other teams included more geometrical complexity in three-dimensional models. The key learning points from Task E have been: • Heat conduction through salt typically requires non-linear thermal conductivity (as a function of temperature), but most models do a good job matching observations, given appropriate adjustments to the applied power and some thermocouple locations. • Thermal pressurization requires coupled thermal-hydrological-mechanical (THM) responses that consider the thermal expansion of the fluid and solid phases. • Initialization of two-phase flow models around a borehole or excavation in salt are more realistically represented as “wetting up”, rather than “drying down” (i.e., the initial state after excavation is mostly dry, rather than mostly wet). • The BATS 1a heater test includes a significant release of brine after the end of heating, which requires a large increase in permeability to recreate. Task E has been a great learning experience for all the teams involved, and feedback from the modeling teams has led to changes in the design of follow-on BATS experiments, which are now ongoing underground at WIPP. There was a balance throughout the task between freedom to model phenomena how each team saw fit, and prescriptiveness in problem design to bring the modeling teams closer together to allow attribution of smaller differences between models to different modeling choices. The modeling approaches seem to go through two phases: an early phase of discovery or testing, and a later phase of refinement and improvement. In future modeling efforts, different field data could be used (e.g., BATS 2) and more time should be included in the processes for teams to make multiple model refinement or even significant changes to their conceptual model or setup, based on lessons learned from the modeling exercise.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

AI Model Benchmarking for Nonproliferation Applications: Steel Thread Benchmarking Task Force Technical Report (Rev. 2)

Steel Thread is a NA-22 venture that seeks to build trustworthy, reliable AI models that can be used in a wide variety of nonproliferation tasks. A key aspect of building these models is developing appropriate benchmarks and evaluation methods, which will enable the venture to identify and adapt models to provide the most value in the nonproliferation domain. Benchmarks must be relevant to key tasks in this domain, such as question answering, information retrieval, document summarization and classification, consensus analysis, and image and data analysis. This report 1) provides an overview of benchmark design, evaluation, and challenges; 2) reviews a variety of open benchmarks, with a focus on language models and tasks; and 3) identifies benchmarks that are most relevant to Steel Thread. This report is intended to serve as a basis for further efforts to classify and evaluate benchmarks and their correlation with success on nonproliferation-specific tasks. The Steel Thread venture has defined benchmarks to be a particular combination of a dataset (or datasets) and a metric (or metrics) conceptualized as representing one or more specific tasks or sets of abilities for a specific modality. It is adopted by a research community as a shared framework for comparing methods.1 It includes 1) Data: Labeled (a designated subset not used for training, which could be all the data), 2) Metric: A way to quantify performance, 3) Task/Ability: What the benchmark is testing, 4) Protocol: A structured and repeatable evaluation process, 5) Baseline/Reference Model: For comparison; could be statistical, rule-based, SME-derived, or another model, and 6) Maintenance Plan: to update with new information over time; important for long-term utility. For further clarity, the definition includes what a benchmark, in this context, is not. It is not a corpus of training data, specific to a model (it is intended to apply to a range of models), a universal evaluation of performance, a guarantee that the ‘top’ model on the leaderboard will be the best fit for every specific use case, an all-encompassing proof of a model’s universal quality, nor is it a one-size-fits-all measure of success. It does not cover every real-world constraint (like operational, ethical, or cost considerations), a systems integration test, or a unit test. This definition was inspired by and resulted from discussions within the Steel Thread Benchmarking Task Force. This group was formed to define what we would mean as a benchmark within Steel Thread but persisted as the need to develop a thorough understanding of the large and expanding existing benchmarking space. This technical report is a result of the group’s divide and conquer approach to exploring this space. The release of benchmarks might not be progressing as quickly as model development, but it is moving very fast, as many benchmarks quickly become saturated, when state-of-the-art models score so close to the benchmark’s ceiling that their results are virtually indistinguishable. At that point, the test no longer differentiates between new systems, so researchers usually stop reporting scores as the benchmark no longer informs about improvements from the next generation of models. In the OpenAI announcement of GPT-5, they reported results on six flagship public benchmarks (AIME 2025, SWE-bench Verified, Aider Polyglot, MMMU, HealthBench Hard, GPQA) but the full system-card covers roughly thirty-five separate evaluations, comprising hundreds of test task items in total. There have been some efforts to summarize benchmarks in specific fields, like for text-to-image generation, but these surveys have had a narrow methodology scope. Therefore, a comprehensive survey of all benchmarks or even all benchmarks that could be relevant to Steel Thread is outside of the scope of this report. We chose some specific benchmarks to investigate in detail.

97 MATHEMATICS AND COMPUTING↗

Personalized learning via task load optimization

A method for providing task load-optimized computer-generated training experiences to a user of a training system that includes: a display, a training simulator, a prediction program (ML1), and a training optimization program (ML2). In response to receiving a predicted optimal task load, ML2 provides a first training experience recommendation related to the training content and/or training conditions that, if utilized in providing a training experience to the user, is predicted to result in the predicted actual task load of the user equaling the predicted optimal task load. In response to receiving biometric information or performance metric information, ML1 determines the predicted actual task load. If the predicted actual task load does not match the predicted optimal task load, ML2 provides a second training experience recommendation and a second training experience is provided where at least one of the training content or the training conditions is changed.

Bertolli, Michael G.↗

Synergistic learning with multi-task DeepONet for efficient PDE problem solving

Multi-task learning (MTL) is an inductive transfer mechanism designed to leverage useful information from multiple tasks to improve generalization performance compared to single-task learning. It has been extensively explored in traditional machine learning to address issues such as data sparsity and overfitting in neural networks. In this work, we apply MTL to problems in science and engineering governed by partial differential equations (PDEs). However, implementing MTL in this context is complex, as it requires task-specific modifications to accommodate various scenarios representing different physical processes. To this end, we present a multi-task deep operator network (MT-DeepONet) to learn solutions across various functional forms of source terms in a PDE and multiple geometries in a single concurrent training session. We introduce modifications in the branch network of the vanilla DeepONet to account for various functional forms of a parameterized coefficient in a PDE. Additionally, we handle parameterized geometries by introducing a binary mask in the branch network and incorporating it into the loss term to improve convergence and generalization to new geometry tasks. Our approach is demonstrated on three benchmark problems: (1) learning different functional forms of the source term in the Fisher equation; (2) learning multiple geometries in a 2D Darcy Flow problem and showcasing better transfer learning capabilities to new geometries; and (3) learning 3D parameterized geometries for a heat transfer problem and demonstrate the ability to predict on new but similar geometries. Finally, our MT-DeepONet framework offers a novel approach to solving PDE problems in engineering and science under a unified umbrella based on synergistic learning that reduces the overall training cost for neural operators.

42 ENGINEERING↗

Asynchronous Execution of Heterogeneous Tasks in ML-Driven HPC Workflows

Heterogeneous scientific workflows consist of numerous types of tasks that require execution on heterogeneous resources. Asynchronous execution of those tasks is crucial to improve resource utilization, task throughput and reduce workflows' makespan. Therefore, middleware capable of scheduling and executing different task types across heterogeneous resources must enable asynchronous execution of tasks. In this paper, we investigate the requirements and properties of the asynchronous task execution of machine learning (ML)-driven high-performance computing (HPC) workflows. We model the degree of asynchronicity permitted for arbitrary workflows and propose key metrics that can be used to determine qualitative benefits when employing asynchronous execution. Our experiments represent relevant scientific drivers, we perform them at scale on Summit, and we show that the performance enhancements due to asynchronous execution are consistent with our model.

97 MATHEMATICS AND COMPUTING↗