Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “tasking”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Asynchronous Execution of Heterogeneous Tasks in ML-Driven HPC Workflows

Heterogeneous scientific workflows consist of numerous types of tasks that require execution on heterogeneous resources. Asynchronous execution of those tasks is crucial to improve resource utilization, task throughput and reduce workflows' makespan. Therefore, middleware capable of scheduling and executing different task types across heterogeneous resources must enable asynchronous execution of tasks. In this paper, we investigate the requirements and properties of the asynchronous task execution of machine learning (ML)-driven high-performance computing (HPC) workflows. We model the degree of asynchronicity permitted for arbitrary workflows and propose key metrics that can be used to determine qualitative benefits when employing asynchronous execution. Our experiments represent relevant scientific drivers, we perform them at scale on Summit, and we show that the performance enhancements due to asynchronous execution are consistent with our model.

97 MATHEMATICS AND COMPUTING↗

Use of machine learning to analyze chemistry card sort tasks

Education researchers are deeply interested in understanding the way students organize their knowledge. Card sort tasks, which require students to group concepts, are one mechanism to infer a student’s organizational strategy. However, the limited resolution of card sort tasks means they necessarily miss some of the nuance in a student’s strategy. Here in this work, we propose new machine learning strategies that leverage a potentially richer source of student thinking: free-form written language justifications associated with student sorts. Using data from a university chemistry card sort task, we use vectorized representations of language and unsupervised learning techniques to generate qualitatively interpretable clusters, which can provide unique insight in how students organize their knowledge. We compared these to machine learning analysis of the students’ sorts themselves. Machine learning-generated clusters revealed different organizational strategies than those built into the task; for example, sorts by difficulty or even discipline. There were also many more categories generated by machine learning for what we would identify as more novice-like sorts and justifications than originally built into the task, suggesting students’ organizational strategies converge when they become more expert-like. Finally, we learned that categories generated by machine learning for students’ justifications did not always match the categories for their sorts, and these cases highlight the need for future research on students’ organizational strategies, both manually and aided by machine learning. In sum, the use of machine learning to analyze results from a card sort task has helped us gain a more nuanced understanding of students’ expertise, and demonstrates a promising tool to add to existing analytic methods for card sorts.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Forecasting for the Weather Driven Energy System - A New Task under IEA Wind

The energy system needs a range of forecast types for its operation in addition to the narrow wind power forecast that has been the focus of considerable recent attention. Therefore, the group behind the former IEA Wind Task 36 Forecasting for Wind Energy has initiated a new IEA Wind Task with a much broader perspective, which includes prospective interaction with other IEA Technology Collaboration Programmes such as the ones for PV, hydropower, system integration, hydrogen etc. In the new IEA Wind Task 51 (entitled "Foreacsting for the Weather Drive Energy System") the existing Work Packages (WPs) are complemented by work streams in a matrix structure. The Task is divided in three WPs according to the stakeholders: WP1 is mainly aimed at meteorologists, providing the weather forecast basis for the power forecasts. In WP2, the forecast service vendors are the main stakeholders, while the end users populate WP3. The new Task 51 started in January 2022. Planned activities include 4 workshops. The first will focus on the state of the art in forecasting for the energy system plus related research issues and be held during September 2022 in Dublin. The other three workshops will be held later during the 4-year Task period and address (1) seasonal forecasting with emphasis on Dunkelflaute, storage and hydro, (2) minute-scale forecasting, and (3) extreme power system events. The issues and conclusions of each of the workshops will be documented by a published paper. Additionally, the Recommended Practice on Forecast Solution Selection will be updated to reflect the broader perspective.

geophysics computing↗

Neck Muscle Coactivation Response to Varied Levels of Mental Workload During Simulated Flight Tasks

Objective To evaluate neck muscle coactivation across different levels of mental workload during simulated flight tasks. Background Neck pain (NP) is highly prevalent among military aviators. Given the complex nature within the flight environment, mental workload may be a risk factor for NP. This may induce higher levels of neck muscle coactivity, which over time may accelerate fatigue, increase neck discomfort, and affect flight task performance. Method Three counterbalanced mental workload conditions represented by simulated flight tasks modulated by interstimulus frequency and complexity were investigated using the Modifiable Multitasking Environment (ModME). The primary measure was a neck coactivation index to describe the neuromuscular effort of the neck muscles as a system. Additional measures included perceived workload (NASA TLX), subjective discomfort, and task performance. Participants ( n = 60; 30M, 30F) performed three test conditions over 1 hr each while seated in a simulated seating environment. Results Neck coactivation indices (CoA) and subjective neck discomfort corresponded with increasing level of mental workload. Average CoAs for low, medium, and high workloads were: .0278(SD = .0232), .0286(SD = .0231), and .0295(SD = .0228), respectively. NASA TLX mental, temporal, effort, and overall scores also increased with the level of mental workload assigned. For ModME task performance, the overall performance score, monitoring accuracy, and resource management accuracy decreased while reaction times increased with the increasing level of mental workload. Communication accuracy was lowest with the low mental workload but had higher reaction times relative to increasing workload. Conclusion Mental workload affects neck muscle coactivation during combinations of simulated flight tasks within a simulated helicopter seating environment. Application The results of this study provide insights into the physical response to mental workload. With increasing multisensory modalities within the work environment, these insights may assist the consideration of physical effects from cognitive factors.

Behavioral Sciences↗

Exploiting Task Tolerances in Mimicry-based Telemanipulation

We explore task tolerances, i.e., allowable position or rotation inaccuracy, as an important resource to facilitate smooth and effective telemanipulation. Task tolerances provide a robot flexibility to generate smooth and feasible motions; however, in teleoperation, this flexibility may make the user’s control less direct. In this work, we implemented a telema nipulation system that allows a robot to autonomously adjust its configuration within task tolerances. We conducted a user study comparing a telemanipulation paradigm that exploits task tolerances (functional mimicry) to a paradigm that requires the robot to exactly mimic its human operator (exact mimicry), and assess how the choice in paradigm shapes user experience and task performance. Our results show that autonomous adjustments within task tolerances can lead to performance improvements without sacrificing perceived control of the robot. Additionally, we find that users perceive the robot to be more under control, predictable, fluent, and trustworthy in functional mimicry than in exact mimicry.

42 ENGINEERING↗

DECOVALEX-2023: Task D Final Report

Task D of DECOVALEX-2023 is focused on the simulation of the coupled thermal hydraulic-mechanical (THM) behaviour in the full-scale engineered barrier system (EBS). The Horonobe EBS experiment is the demonstration of the full-scale EBS in the underground research laboratory (URL) (performed by JAEA in the Horonobe URL in Japan). Task D consisted of the three steps, a preliminary step (Step 0), simulation of the laboratory tests (Step 1) and simulation of the in-situ full-scale EBS experiment (Step 2). Since the Horonobe EBS experiment demonstrates the vertical emplacement option of the EBS, the experiment gallery is also backfilled with the backfill material. Therefore, interaction between the EBS and the backfill material can also be demonstrated, such as deformation (change of density) of the buffer material. The underground water in the Horonobe URL is saline. This fact adds chemical processes to THM behaviour. For example, mechanical properties (such as swelling pressure of the buffer material and backfill material) and hydraulic properties (such as permeability of the buffer material and backfill material) change depending on the water chemistry. Task D was therefore a challenging Task focused on not only the relatively simple THM behaviour but also complex THM behaviour including chemical processes. Six research teams (BGR, CAS, JAEA, KAERI, SNL and Taipower) participated the Task D. BGR, CAS, JAEA, KAERI and Taipower research teams selected a THM approach, while the SNL research team selected a TH approach. Step 1 involved the simulation of laboratory test results and was important to check the numerical codes developed by the research teams. Step 1 was divided into four sub steps. The simulation results through the Step 1 identified the parameters for simulation of the Step 2. Basic parameters of the materials (buffer material, backfill material, rock mass, concrete, sand) were provided by JAEA. Special parameters which research team needed were identified by back analysis of Step 1. Most notably the mechanical behaviour of swelling and displacement depended on the applied model (elastic model or elastoplastic model). Parameters such as Young’s modulus were found to need smaller values than characterised in the fundamental laboratory test results (Step 1-1, 1-2) for the elastic model. Although laboratory experiments are usually simple, test results contained some error. For example, if the saturation level is 100 % or higher, it should be considered an error. This situation was presented in the Step 1-3. A possible reason is that the buffer material is a mixture of bentonite and silica sand. When a specimen is cut to measure volume or weight, sand grains will affect the measurement data. In Step 2, boundary conditions such as temperature on the surface of the simulated overpack, heater power of the electrical heaters installed in the simulated overpack, injection pressure and inflow rate of the test water, were applied. The outer boundary conditions can be selected using measured data (injection pressure and inflow rate of the test water that is controlled by the injection systems installed in the sand layer around the buffer material and in the boundary between backfill material and concrete support). Since such measured data has some noise, research teams developed their own simplified boundary conditions. Inner boundary conditions can be selected using measured data as heater power and temperature on the surface of the simulated overpack. These data also contain some noise, so research teams developed their own simplified developed boundary conditions. Task D validated various approaches thorough the simulation of the in-situ full scale EBS system including backfill of the gallery: variations in the coupling processes (THM or THC), analysis codes, and boundary conditions. Temperature distribution in the buffer material was simulated well by all research teams. This means thermal behaviour is not sensitive to the simulation approaches. Although the water content distribution on the outside of the buffer material was well simulated by all research teams, the simulation results differ from the measured values inside the buffer material (at the centre and inside, near the simulated overpack). The buffer material is made from tap water, but in the in-situ experiment, saline groundwater infiltrates the buffer material. Therefore, the selection of the hydraulic parameters of the buffer material greatly affects the simulation results of the re saturation behaviour of the buffer material. In the Horonobe EBS experiment, measured values suitable for validating the simulation results were not obtained near the simulated overpack. When simulating the pressure and deformation of the buffer material, the measurement data is easily affected by the installation conditions of the measurement sensors, so verifying the measurement data itself remains an issue. Mechanical simulation results differ depending on whether they are considered as elastic or elastoplastic phenomena. The accuracy of measured in-situ data can be assessed by detailed analysis comparing sampling specimen analysis and measured data. The Horonobe EBS experiment is scheduled to be dismantled in the future (FY2026 and 2027). This detailed dismantling investigation will finally confirm the measured data.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

Evaluation of Howard A. Hanson Dam Juvenile Fish Passage and Survival Study Live Fish Injury Assessment, Sensor Fish, and BioPA Modeling Tasks

The live fish injury assessment, Sensor Fish, and BioPA modeling study tasks were conducted by researchers from Pacific Northwest National Laboratory (PNNL). The four tasks were part of the larger Evaluation of Howard A. Hanson Dam (HAHD) Juvenile Fish Passage and Survival study, which had six total tasks. To achieve study objectives for each of the four tasks, field work occurred at Green Peter Dam (GPR) to evaluate the highest elevation steep slope bypass pipe, at HAHD to evaluate baseline conditions of the horseshoe tunnel, and at PNNL’s Aquatic Research Laboratory (ARL) to evaluate simulated dam passage conditions (i.e., shear forces and collision). Each of these evaluations utilized live fish injury assessment, Sensor Fish, and BioPA modeling. Live fish injury assessment and survival (tagged with and without balloon or passive integrated transponder [PIT] tags) was correlated with Sensor Fish to determine thresholds. The CFD analyses were then performed, and the computed values were compared to the corresponding measured values of Sensor Fish data. The results of the overall injury and survival of fish was also used in the validation of the CFD modeling method. Collectively, the results will aid in future modeling of fish passage at HAHD. Results from these tasks can be used by biologists, engineers, resource managers, and regional decision-makers to inform baseline conditions under current operations and the engineering design of the new FPF at HAHD. This draft report contains initial data and results from the four tasks. Table 8 1, Table 8 2, and Table 8 3, and Figure 8 1, Figure 8 2, and Figure 8 3 depict the CFD modeling findings for the GPR steep slope bypass, HAHD horseshoe tunnel, and laboratory testing. Table 8 4, Table 8 5, and Table 8 6 depict the Sensor Fish findings for the GPR steep slope bypass and HAHD horseshoe tunnel testing. The Mv values observed in the HAHD were significantly lower compared to the laboratory experiments conducted at PNNL. Currently, investigations are underway to understand the reasons for this disparity and to establish an appropriate threshold value for Mv. Survival predictions presented in the tables below should be considered preliminary and should not be used until further analyses and adjustments are completed. The next steps for modeling will include the flow regime, (i.e., density of flow regimes due to water and air mixing ) to continue to improve on the threshold value for Mv.

13 HYDRO ENERGY↗

Large Load Integration - Task List and Overview

Large Load Integration Tasks: Task 1 – Workshops Support stakeholder engagement across industry to promote collaboration and identify solutions to challenges that will guide other work Task 2 – Ancillary Services Characterize different types of large loads to assess under what conditions they may be utilized to provide grid stability services Task 3 – Communications Explore the cybersecurity and communications infrastructure required to enable large loads to interface with grid operations to provide ancillary services Task 4 – Nuclear Integration Explore risks and methods for supporting large load energy needs with SMRs and incorporating them into the wider power system Task 5 – Decision Support and TA Provide support to stakeholders through the creation of planning tools and direct technical assistance.

24 - POWER TRANSMISSION AND DISTRIBUTION↗

Performance Characterization and Provenance of Distributed Task-based Workflows on HPC Platforms

Understanding performance and provenance of task-based workflows poses significant challenges, particularly in distributed configurations where resources are shared by multiple applications. Task-based workflow management systems further complicate performance predictability because of their dynamicity that subtly alters task execution order from run to run. In this paper we propose a layered characterization framework for performance and task provenance for Dask.distributed workflows running on high-performance computing (HPC) platforms. It collects data from jobs, the workflow management system, and the operating system to aid in understanding the performance of these workflows. Our approach encompasses three main contributions: first, an extension of Dask.distributed to capture high-fidelity task provenance using Mochi data services; second, the adaptation of the established HPC I/O characterization tool Darshan to gather high-fidelity I/O data, thereby enhancing the granularity of our analysis; and third, a framework to combine and process the collected data and provide helpful insights into performance characterization and reproducibility, alongside our lessons learned.

Dask↗

A Hierarchical Task Scheduler for Heterogeneous Computing

Heterogeneous computing is one of the future directions of HPC. Task scheduling in heterogeneous computing must balance the challenge of optimizing the application performance and the need for an intuitive interface with the programming run-time to maintain programming portability. The challenge is further compounded by the varying data communication time between tasks. This paper proposes RANGER, a hardware-assisted task-scheduling framework. By integrating RISC-V cores with accelerators, the RANGER scheduling framework divides scheduling into global and local levels. At the local level, RANGER further partitions each task into fine-grained subtasks to reduce the overall makespan. At the global level, RANGER maintains the coarse granularity of the task specification, thereby maintaining programming portability. The extensive experimental results demonstrate that RANGER achieves a 12.7× performance improvement on average, while only requires 2.7% of area overhead.

Miniskar, Narasinga Rao↗

Regional specialization in prefrontal cortex manifests in the reliability of task progression codes

The brain has the remarkable ability to guide the performance of complex tasks. Distinct prefrontal cortical areas make specific contributions to this ability, with the orbitofrontal cortex (OFC) critical for processing information related to trial outcomes and the dorsomedial prefrontal cortex (dmPFC) critical for sustained effort and selecting the right action at the right time. Yet, in both areas, neural activity represents both outcome- and action-related quantities. How similar neural representations support different functions remains unclear. Here, we compared OFC and dmPFC activity in rats performing a spatial alternation task. We show that, in contrast to other task-related variables, task progression is represented in both areas, but with distinct patterns of across-trial reliability that match each area’s previously documented functional specialization. Our results indicate that the engagement of reliable, task-phase-specific activity patterns differs across prefrontal regions in a manner well suited to engage different computations at different times.

Biological and medical sciences↗

GRUMDN: A Multi-Task Model for Predicting Human Patterns-of-Life from Stay Transition Data

Understanding human patterns-of-life (PoL) is essential towards ensuring safe and secure indoor facility environment as well as outdoor urban environment. Prediction of human movement in between places of interest is vital in understanding human PoL. Movement between spaces maybe represented and detected in one of the two forms: 1) trajectories: locations measured at regular time intervals by mobile sensors, bluetooth or GPS sensors; or 2) stay transitions: semantic PoI (points of interest) and stay duration data measurable by eventbased sensors that collect data when a check-in or check-out event is detected. Stay transition data provides a more compressed data format compared to trajectories data, especially in situations with longer stay durations, while preserving the information necessary for PoL analysis. Now as introduced briefly in the paper, our deployed end application (Digital Twin of a facility with non-player characters, besides the interactive user in virtual reality) needed a well-performing and validated AI/ML model for simulating high quality stay transitions behavior. In this study we thus primarily present our findings with developing and validating that model, which is a multi-task neural network for stay transition prediction. The neural network consists of two heads, for corresponding two tasks of stay category prediction and stay duration prediction. We evaluated gated recurrent units and multi-layer perceptrons of varying network sizes for stay category prediction; while mixture density networks, noisy generator-only networks, and generative adversarial networks of varying network sizes for stay duration prediction. We have then evaluated four multi-task models, constructed by combining these specialized models, on their ability to predict stay transition data. We tested our models on datasets from two different cases: 1) a simulation-generated dataset of indoor movement within the HFIR (high flux isotope reactor) nuclear reactor facility at Oak Ridge National Laboratory (ORNL); and 2) the GeoLife human mobility dataset of outdoor urban movement available in literature. Our results indicate that GRUMDN, which combines gated recurrent units (GRU) for stay category prediction task, and mixture density networks (MDN) for stay duration prediction task, did overall outperform other multitask models and the current state-of-the-art.

Gunaratne, Chathika [ORNL] (ORCID:0000000225088745↗

The impacts of training on change deafness and build-up in a flicker task

Performance on auditory change detection tasks can be improved by training. We examined the stimulus specificity of these training effects in behavior and ERPs. A flicker change detection task was employed in which spatialized auditory scenes were alternated until a "change" or "same" response was made. For half of the trials, scenes were identical. The other half contained changes in the spatial locations of objects from scene to scene. On Day 1, participants were either trained on this auditory change detection task (trained group), or trained on a non-auditory change detection task (control group). On Day 2, all participants were tested on the flicker task while EEG was recorded. The trained group showed greater change detection accuracy than the control group. They were less biased to respond "same" and showed full generalization of learning from trained to novel auditory objects. ERPs for "change" compared to "same" trials showed more negative going P1, N1, and P2 amplitudes, as well as a larger P3b amplitude. The P3b amplitude also differed between the trained and control group, with larger amplitudes for the trained group. Analysis of ERPs to scenes viewed prior to a decision revealed build-up of a difference between "change" and "same" trials in N1 and P2. Results demonstrate that training has an impact early in the "same" versus "change" decision-making process, and that the flicker paradigm combined with the ERP method can be used to study the build-up of change detection in auditory scenes.

60 APPLIED LIFE SCIENCES↗

DECOVALEX-2023 Task F Specification (Rev. 9)

This report is the revised (Revision 9) Task F specification for DECOVALEX-2023. Task F is a comparison of the models and methods used in deep geologic repository performance assessment. The task proposes to develop a reference case for a mined repository in a fractured crystalline host rock (Task F1) and a reference case for a mined repository in a salt formation (Task F2). Teams may choose to participate in the comparison for either or both reference cases. For each reference case, a common set of conceptual models and parameters describing features, events, and processes that impact performance will be given, and teams will be responsible for determining how best to implement and couple the models. The comparison will be conducted in stages, beginning with a comparison of key outputs of individual process models, followed by a comparison of a single deterministic simulation of the full reference case, and moving on to uncertainty propagation and uncertainty and sensitivity analysis. This report provides background information, a summary of the proposed reference cases, and a staged plan for the analysis.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

DECOVALEX-2023: Task F Specification (Revision 10)

This report is the revised (Revision 10) Task F specification for DECOVALEX-2023. Task F is a comparison of the models and methods used in deep geologic repository performance assessment. The task proposes to develop a reference case for a mined repository in a fractured crystalline host rock (Task F1) and a reference case for a mined repository in a salt formation (Task F2). Teams may choose to participate in the comparison for either or both reference cases. For each reference case, a common set of conceptual models and parameters describing features, events, and processes that impact performance will be given, and teams will be responsible for determining how best to implement and couple the models. The comparison will be conducted in stages, beginning with a comparison of key outputs of individual process models, followed by a comparison of a single deterministic simulation of the full reference case, and moving on to uncertainty propagation and uncertainty and sensitivity analysis. This report provides background information, a summary of the proposed reference cases, and a staged plan for the analysis.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

DECOVALEX-2023: Task A Final Report

Task A, also known as HGFrac, examines the fracturing processes that may occur in the Callovo-Oxfordian claystone (COx) in the context of the high-level (HLW) and intermediate-level long-lived (ILW-LL) radioactive waste repository in France. Understanding these processes and improving numerical models to reproduce them will aid in the design, optimization, and safety of the repository. Heat and gas fracturing are studied in two independent subtasks following a stepwise approach: laboratory tests/benchmark exercises, in-situ experiment, and finally, an application case. The in-situ heater experiment aimed to thermally induce a hydraulic fracture; temperature and pore pressure were monitored to detect any evidence of fracturing. The in-situ gas injection experiment aimed to study the effect of the stress orientation and gas injection kinetics on the gas fracturing process. The occurrence of fracturing was monitored by gas pressure measured in the injection interval. In both experiments, the excavation-induced fracture network around the heater/injection boreholes played an important role in the reduction of the compressive stress state, leading to both a tensile and shear failure response of the COx. During the first two years of the project, the research teams working on each task developed and/or proposed numerical approaches for reproducing the occurrence of fracturing in the in-situ experiments. The failure criteria were defined by reproducing the measurements from laboratory extension tests for the heat fracturing subtask. In the gas fracturing subtask, their approaches were used for simulating several benchmark exercises, and an inter-comparison between models was carried out. In both tasks, the developed approaches were compared with a simplified approach considering poro-elasticity for the mechanical behaviour of the COx. Most of the developed approaches are based on a continuous medium that takes into account variations in hydraulic properties due to mechanical degradation, such as plastic deformation or damage. Other approaches implicitly modelled weak planes or embedded discontinuities to reproduce fracture propagation. The potential for fracture initiation was also studied through of a discrete approach. In the second half of the project, the research teams mainly focused on interpretative modelling of two in-situ experiments and a blind prediction exercise to test their respective approaches. The models developed by the research teams were also applied at the repository scale to evaluate fracture initiation in a case study under vi unfavourable conditions, particularly in terms of spacing between High-Level Waste cells. The results showed that the poro-elasticity approach could be an efficient tool for understanding the main processes occurring in the COx. One example is the explicit representation of the excavation-induced fracture network around the boreholes, which yielded acceptable results compared to the measurement data. However, advanced approaches were needed to evaluate the potential increase of the excavation-induced fracture network extend and better understand fracture initiation. The stress analyses carried out by the teams revealed that hydraulic boundary conditions had a strong impact on fracture initiation in the heater experiment. Furthermore, in most cases, the results required higher pore pressure increments to reach fracturing than those measured in the experiment. This implies that the measurements may have been biased by the packer’s capacity to fully isolate the piezometric chambers, leading to lower pressures. On the contrary, there was no agreement on the fracturing mode, as some reported either shear or tensile fracturing, while others reported a combination of the two modes. In the case of gas fracturing, the research teams were limited to the comparison of a single point, which complicated their task. Nonetheless, the numerical results were able to reproduce the measurements and capture processes such as longitudinal gas flow through the excavation-induced fracture network, as suggested by some evidence in the observation piezometric chambers. The numerical models also agreed with the measurements in the sense of higher probability of developing along the injection borehole than radially towards the sound rock. The approaches developed by the research teams showed that they are capable of analysing and reproducing fracture initiation in the COx. However, areas of future work should focus on the fracture propagation and fracture aperture, which were out of the scope of this task. To this end, additional data must be gathered for the parameter characterisation and validation of the numerical models. Nonetheless, various approaches showed promising results as they were able to reproduce fracture development under certain conditions.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

Overview of the 2021 SDP 3C Citation Context Classification Shared Task

This paper provides an overview of the 2021 3C Citation Context Classification shared task. The second edition of the shared task was organised as part of the 2nd Workshop on Scholarly Document Processing (SDP 2021). The task is composed of two subtasks: classifying citations based on their (Subtask A) purpose and (Subtask B) influence. As in the previous year, both tasks were hosted on Kaggle and used a portion of the new ACT dataset. A total of 22 teams participated in Subtask A, and 19 teams competed in Subtask B. All the participated systems were ranked based on their achieved macro f-score. The highest scores of 0.26973 and 0.60025 were reported for sub-task A and B, respectively.

Kunnath, Suchetha N.↗

Multi-task graph neural networks for simultaneous prediction of global and atomic properties in ferromagnetic systems *

Abstract We introduce a multi-tasking graph convolutional neural network, HydraGNN, to simultaneously predict both global and atomic physical properties and demonstrate with ferromagnetic materials. We train HydraGNN on an open-source ab initio density functional theory (DFT) dataset for iron-platinum with a fixed body centered tetragonal lattice structure and fixed volume to simultaneously predict the mixing enthalpy (a global feature of the system), the atomic charge transfer, and the atomic magnetic moment across configurations that span the entire compositional range. By taking advantage of underlying physical correlations between material properties, multi-task learning (MTL) with HydraGNN provides effective training even with modest amounts of data. Moreover, this is achieved with just one architecture instead of three, as required by single-task learning (STL). The first convolutional layers of the HydraGNN architecture are shared by all learning tasks and extract features common to all material properties. The following layers discriminate the features of the different properties, the results of which are fed to the separate heads of the final layer to produce predictions. Numerical results show that HydraGNN effectively captures the relation between the configurational entropy and the material properties over the entire compositional range. Overall, the accuracy of simultaneous MTL predictions is comparable to the accuracy of the STL predictions. In addition, the computational cost of training HydraGNN for MTL is much lower than the original DFT calculations and also lower than training separate STL models for each property.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗