Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “parallel time integration”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 469 records · Page 26

Extending Validated Human Performance Models to Explore NextGen Concepts

To meet the expected increases in air traffic demands, NASA and FAA are researching and developing Next Generation Air Transportation System (NextGen) concepts. NextGen will require substantial increases in the data available to pilots on the flight deck (e.g., weather,wake, traffic trajectory predictions, etc.) to support more precise and closely coordinated operations (e.g., self-separation, RNAV/RNP, and closely spaced parallel operations, CSPOs). These NextGen procedures and operations, along with the pilot's roles and responsibilities, must be designed with consideration of the pilot's capabilities and limitations. Failure to do so will leave the pilots, and thus the entire aviation system, vulnerable to error. A validated Man-machine Integration and design Analysis System (MIDAS) v5 model was extended to evaluate anticipated changes to flight deck and controller roles and responsibilities in NextGen approach and Land operations. Compared to conditions when the controllers are responsible for separation on decent to land phase of flight, the output from these model predictions suggest that the flight deck response time to detect the lead aircraft blunder will decrease, pilot scans to the navigation display will increase, and workload will increase.

Gore, Brian Francis↗

Benchmarking of massively parallel phase-field codes for directional solidification

We present a detailed benchmark comparing two state-of-the-art phase-field implementations for simulating alloy solidification under experimentally relevant conditions. The study investigates the directional solidification of Al-3wt%Cu under high-velocity solidification conditions and SCN-0.46wt% camphor under microgravity conditions from National Aeronautics and Space Administration (NASA) DECLIC-DSI-R experiments. Both codes, one employing finite-difference discretization with uniform mesh and GPU-acceleration (GPU-PF) and the other one employing finite-element discretization with adaptive-mesh and CPU-parallelization (PRISMS-PF), solve the same quantitative phase-field formulation that incorporates an anti-trapping current for the solidification of dilute alloys. We evaluate the predictions of each code for dendritic morphology, primary spacing, and tip dynamics in both 2D and 3D, as well as their numerical convergence and computational performance. While existing benchmark problems have primarily focused on simplified or small-scale simulations, they do not reflect the computational and modeling challenges posed by employing experimentally relevant time and length scales. Our results provide a practical framework for assessing phase-field code performance as well as validating and facilitating their application in integrated computational materials engineering (ICME) workflows that require integration with realistic experimental data.

36 MATERIALS SCIENCE↗

Probabilistic Analysis of Impact of Wake Vortices on Closely-Spaced Parallel Approaches

One of the primary constraints on the capacity of the nation's air transportation system is the landing capacity of its largest airports. Many airports with closely spaced parallel runways suffer a severe runway acceptance rate when the weather conditions do not allow full utilization of these parallel runways. The present requirement for simultaneous independent landings in Instrument Meteorological Conditions, IMC, is at least 4300 feet of lateral runway spacing (as close as 3000 feet for runways with a Precision Runway Monitor). Operations in Visual Meteorological Conditions, VMC, to Closely Spaced Parallel Approaches only require a lateral runway spacing greater than 750 feet. A study by Hardy and Lewis integrated and extended earlier studies and concepts in lateral traffic separation, longitudinal station keeping, wake prediction, wake display, and the concepts of R N P into a preliminary system concept for Closely Spaced Parallel Approaches in IMC. This system allows IMC airport acceptance rates to approach those for VMC. The system concept that was developed, presented traffic and wake information on the NAVigation Display, NAV, and developed operational procedures for a mix of conventional and Runway Independent Aircraft with different approach speeds to Closely Spaced Parallel Runways. This paper first describes some improvements made on the technology needed to better predict and formulate a probabilistic representation for the time-dependent motion and spreading of the hazardous region associated with the lift-generated vortex wakes of preceding aircraft. In this way, the time at which the vortex wakes of leading aircraft intrude into the airspace of adjacent flight-corridor/runway combinations can be more reliably predicted. Such a prediction is needed because it determines restraints to be placed on in-trail separation distances; or, the allowable time intervals between aircraft executing nearly simultaneous landings or takeoffs on very closely-spaced runways. Improved estimates of wake spreading are achieved by inclusion of representations in the equations for wake spreading due to ambient turbulence and due to the long-wave instability of a vortex pair. Wake motion and spreading due to the time-averaged wind and its variations with time, are retained. The more detailed representation of wake spreading presented here permits the development of probabilistically-based uncertainty estimates for wake spreading. Measurements needed within actual aircraft wake vortices to validate and support this analysis are also described. The second part of the paper uses the improvements in the accuracy of the location of wake vortices to extend the preliminary system concept for Closely Spaced Parallel Approaches described earlier with more robust operational procedures. Additionally, improvements in longitudinal station keeping, wake display, and risk assessment methodologies are incorporated and described.

Hardy, Gordon H.↗

Graph-Learning-Assisted State and Event Tracking for Solar-Penetrated Power Grids with Heterogeneous Data Sources

Unlike transmission systems, distribution systems do not typically contain sufficient metering to enable real-time state estimation. The lack of sufficient real-time measurements prohibits accurate and timely monitoring of the state of distribution systems. As a result, control and optimal operation of distribution systems, especially those containing large numbers of renewable generation units are not possible without proper data and information about the current state of the system. The main motivation of this project is to address this shortcoming by developing an approach which provides “predicted” real-time measurements so that they can be used to execute a distribution system state estimator. Thus, the objective of the project is to make the distribution systems fully observable, such that the hosting capacity for solar generation can be accurately estimated, and unnecessary solar curtailments can be avoided. In order to accomplish this goal, the project investigated the use of a grid-model-informed machine learning (ML) tool which integrates heterogeneous data streams obtained from AMI meters, SCADA as well as PMU measurements and created synchronous measurement snapshots for the state estimator (SE); and developed a hybrid robust SE which provides not only accurate state estimates but also real-time feedback for the ML model refinement.

14 SOLAR ENERGY↗

Parallel Subconvolution Filtering Architectures

These architectures are based on methods of vector processing and the discrete-Fourier-transform/inverse-discrete- Fourier-transform (DFT-IDFT) overlap-and-save method, combined with time-block separation of digital filters into frequency-domain subfilters implemented by use of sub-convolutions. The parallel-processing method implemented in these architectures enables the use of relatively small DFT-IDFT pairs, while filter tap lengths are theoretically unlimited. The size of a DFT-IDFT pair is determined by the desired reduction in processing rate, rather than on the order of the filter that one seeks to implement. The emphasis in this report is on those aspects of the underlying theory and design rules that promote computational efficiency, parallel processing at reduced data rates, and simplification of the designs of very-large-scale integrated (VLSI) circuits needed to implement high-order filters and correlators.

Gray, Andrew A.↗

Assessment methods for determining small changes in hearing performance over time

Although the behavioral pure-tone threshold audiogram is considered the gold standard for quantifying hearing loss, assessment of speech understanding, especially in noise, is more relevant to quality of life but is only partly related to the audiogram. Metrics of speech understanding in noise are therefore an attractive target for assessing hearing over time. However, speech-in-noise assessments have more potential sources of variability than pure-tone threshold measures, making it a challenge to obtain results reliable enough to detect small changes in performance. Here, this review examines the benefits and limitations of speech-understanding metrics and their application to longitudinal hearing assessment, and identifies potential sources of variability, including learning effects, differences in item difficulty, and between- and within-individual variations in effort and motivation. We conclude by recommending the integration of non-speech auditory tests, which provide information about aspects of auditory health that have reduced variability and fewer central influences than speech tests, in parallel with the traditional audiogram and speech-based assessments.

60 APPLIED LIFE SCIENCES↗

Development and Validation of a Fast, Accurate and Cost-Effective Aeroservoelastic Method on Advanced Parallel Computing Systems

Progress to date towards the development and validation of a fast, accurate and cost-effective aeroelastic method for advanced parallel computing platforms such as the IBM SP2 and the SGI Origin 2000 is presented in this paper. The ENSAERO code, developed at the NASA-Ames Research Center has been selected for this effort. The code allows for the computation of aeroelastic responses by simultaneously integrating the Euler or Navier-Stokes equations and the modal structural equations of motion. To assess the computational performance and accuracy of the ENSAERO code, this paper reports the results of the Navier-Stokes simulations of the transonic flow over a flexible aeroelastic wing body configuration. In addition, a forced harmonic oscillation analysis in the frequency domain and an analysis in the time domain are done on a wing undergoing a rigid pitch and plunge motion. Finally, to demonstrate the ENSAERO flutter-analysis capability, aeroelastic Euler and Navier-Stokes computations on an L-1011 wind tunnel model including pylon, nacelle and empennage are underway. All computational solutions are compared with experimental data to assess the level of accuracy of ENSAERO. As the computations described above are performed, a meticulous log of computational performance in terms of wall clock time, execution speed, memory and disk storage is kept. Code scalability is also demonstrated by studying the impact of varying the number of processors on computational performance on the IBM SP2 and the Origin 2000 systems.

Goodwin, Sabine A.↗

Modernization of the Radiation Measurements Laboratory at the Advanced Test Reactor Complex

The Advanced Test Reactor (ATR) at the Idaho National Laboratory (INL) is a unique, water-cooled, high-flux test reactor capable of performing tests prototypical of PWR operating conditions. The Radiation Measurements Laboratory (RML) was founded in the 1960s to support reactor operations and to conduct independent scientific research. For nearly six decades RML has performed four primary functions: monitoring of radioactivity by gamma-ray spectroscopy of routine reactor samples, fluence rate determinations for irradiation cycles, fission-rate measurements for the ATR-Critical (ATR-C) Facility, and independent research and development of radiation detection systems and applications. Through the decades, RML has seen technological advancements that have been integrated into each of the critical functions of the laboratory. However, many of the measurement and analysis systems employed to this day can be improved by modernization. Recent progress at RML is improving reliability and accuracy of the radiation measurements performed in support of nuclear energy research for the U.S.A. Department of Energy. The control and data collection systems supporting ATR-C have been upgraded. New High-Purity Germanium (HPGe) spectrometers have been procured with liquid nitrogen recycling capabilities to improve up-time and reduce measurement uncertainties. Fluence-rate measurement techniques are also being improved to ensure accuracy, avoid systemic errors, and identify biases. In a parallel effort, new scientific research avenues are being explored which will provide an opportunity to further enhance the utilization of ATR and ensure the sustainability of the RML as nuclear research continues to evolve. The RML is improving the effectiveness of irradiation services provided by ATR while ensuring a sustainable future for nuclear energy research.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Multi-Source Machine Learning and Thermoplastics Enhanced Aerostructure Manufacturing (mTEAM)

RTX Technology Research Center (RTRC), together with Collins Aerospace (Collins) and Oak Ridge National Laboratory (ORNL) has developed an Artificial Intelligence (AI) / Machine Learning (ML) guided solution to advance the manufacturing and assembly of high performance and lightweight thermoplastic composite (TPC) aerospace products. The solution aims to lower risk, cost and lead time for induction heating based welding and consolidation processes for TPC structure. The cost and lead time of part and material specific process development for induction welding (IW) and induction consolidation will be reduced by replacing traditional empirical methods with optimization methods that merge AI/ML and physics-based process simulations and process experiments with sensing and controls. TPC-IW process development is empirical in nature, and uncertainties in material & process behavior exist near & far from the induction coil. Physics-based simulations can be leveraged directly for process optimization but can be too computationally expensive to run in high fidelity and real time to do robust process optimization. The key impact of successful TPC induction consolidation and welding is cost & lead time reduction for part & material specific consolidation and welding recipes. This is an enabler for more rapid deployment of TPC structures via joining assembly, which can reduce energy & cost intensive usage of autoclaves & ovens. The solution aimed to advance the U.S. Department of Energy’s interests in using thermoplastics and automation in composite manufacturing for improvement of products for existing markets via increased production speeds, reduced costs, and lowered use of energy. Welded TPC structures can offer significant weight & energy savings for high-value commercial aerospace & industrial applications compared to metal & thermoset composite structures assembled by mechanical fastening and/or adhesive bonding. The project was organized into two Budget Periods. Budget Period 1 (BP1) was 15 months and its goal was to perform ML process optimization framework development & deployment on lab-coupon aerostructure components. A Go/No-Go Review was performed at the end of BP1 to verify fulfilment of key tasks & milestones to justify a Go Decision to move into the next Budget Period. Budget Period 2 (BP2) was 12 months and its goal was the deployment of the ML framework for ML process optimization of pilot industrial scale aerostructure components. The overall project aim was to develop & demonstrate ML-enhanced modeling framework that learns process-property mapping from multiple data sources at different fidelities. During BP1, the team accomplished key tasks & milestones to demonstrate the concept of multi-source ML for TPC aerostructure consolidation and assembly. First, the team completed documentation of induction based TPC heating requirements including baseline metrics to compare measured results against. Next the team completed demonstration of data generation from physics-based simulations for ML surrogate model generation and demonstrated the integration of physics-based simulation data into multi-source AI/ML algorithms. In parallel, the team established the lab-coupon scale induction welding system and completed a process to label and reduce generated data from physics-based simulation and experiments for ML surrogate models to enable multi-source ML model training & testing. To complete BP1, the team integrated physics-based simulation data and experimental data into multi-source ML algorithms. This was based on the team completing ML deployment of the induction welding on a lab system at RTRC and AI/ML deployment on existing induction welding line at Collins. ORNL visited both Collins and RTRC sites to witness the TPC induction welding process. Then, ORNL designed and constructed a new version of their vision-based sensing system better adapted to acquire process signals of the TPC induction welding process for process anomaly and defect detection. In BP2, the team accomplished key tasks & milestones to scale up multi-source ML for TPC aerostructure consolidation and assembly from the lab-coupon scale to the pilot-industrial scale. In BP2, the team demonstrated real time anomaly & defect detection via experiments performed by ORNL & RTRC. The team completed ML-optimization heating trials for TPC induction consolidation at Collins, and the team confirmed pilot industrial scale experimental data from Collins was compatible with the developed ML pipeline from RTRC. The team completed sub-element scale ML process optimization demonstration at RTRC, where the team leveraged RTRC’s robotic TPC welding setup to de-risk the ML process optimization by performing ML analysis of recorded temperatures to account for complex part features. Then, the team applied its ML-derived control strategies and ML process optimization framework at Collins to the pilot-industrial scale on a demo skin-stiffener part representative of a nacelle aerostructure fan cowl section. The key innovation is the AI/ML framework enabling effective process development of high performance, lightweight, energy efficient TPCs for composite aircraft structures.

36 MATERIALS SCIENCE↗

Development of a Robust and Efficient Parallel Solver for Unsteady Turbomachinery Flows

The traditional design and analysis practice for advanced propulsion systems relies heavily on expensive full-scale prototype development and testing. Over the past decade, use of high-fidelity analysis and design tools such as CFD early in the product development cycle has been identified as one way to alleviate testing costs and to develop these devices better, faster and cheaper. In the design of advanced propulsion systems, CFD plays a major role in defining the required performance over the entire flight regime, as well as in testing the sensitivity of the design to the different modes of operation. Increased emphasis is being placed on developing and applying CFD models to simulate the flow field environments and performance of advanced propulsion systems. This necessitates the development of next generation computational tools which can be used effectively and reliably in a design environment. The turbomachinery simulation capability presented here is being developed in a computational tool called Loci-STREAM [1]. It integrates proven numerical methods for generalized grids and state-of-the-art physical models in a novel rule-based programming framework called Loci [2] which allows: (a) seamless integration of multidisciplinary physics in a unified manner, and (b) automatic handling of massively parallel computing. The objective is to be able to routinely simulate problems involving complex geometries requiring large unstructured grids and complex multidisciplinary physics. An immediate application of interest is simulation of unsteady flows in rocket turbopumps, particularly in cryogenic liquid rocket engines. The key components of the overall methodology presented in this paper are the following: (a) high fidelity unsteady simulation capability based on Detached Eddy Simulation (DES) in conjunction with second-order temporal discretization, (b) compliance with Geometric Conservation Law (GCL) in order to maintain conservative property on moving meshes for second-order time-stepping scheme, (c) a novel cloud-of-points interpolation method (based on a fast parallel kd-tree search algorithm) for interfaces between turbomachinery components in relative motion which is demonstrated to be highly scalable, and (d) demonstrated accuracy and parallel scalability on large grids (approx 250 million cells) in full turbomachinery geometries.

West, Jeff↗

Evaluating the De Hoffmann-Teller Cross-Shock 2 Potential at Real Collisionless Shocks

Shock waves are common in the heliosphere and beyond. The collisionless nature of most astrophysical plasmas allows for the energy processed by shocks to be partitioned amongst particle sub-populations and electromagnetic fields via physical mechanisms that are not well understood. The electrostatic potential across such shocks is frame dependent. In a frame where the incident bulk velocity is parallel to the magnetic field, the deHoffmann-Teller frame, the potential is linked directly to the ambipolar electric field established by the electron pressure gradient. Thus measuring and understanding this potential solves the electron partition problem, and gives insight into other competing shock processes. Integrating measured electric fields in space is problematic since the measurements can have offsets that change with plasma conditions. The offsets, once integrated, can be as large or larger than the shock potential. Here we exploit the high-quality field and plasma measurements from NASA’s Magnetospheric Multiscale mission to attempt this calculation. We investigate recent adaptations of the deHoffmann-Teller frame transformation to include time variability, and conclude that in practice these face difficulties inherent in the 3D time-dependent nature of real shocks by comparison to 1D simulations. Potential estimates based on electron fluid and kinetic analyses provide the most robust measures of the deHoffmann-Teller potential, but with some care direct integration of the electric fields can be made to agree. These results suggest that it will be difficult to independently assess the role of other processes, such as scattering by shock turbulence, in accounting for the electron heating.

Steven J Schwartz↗

ROOT RNTuple and EOS: The Next Generation of Event Data I/O

For several years, the ROOT team is developing the new RNTuple I/O subsystem in preparation of the next generation of collider experiments. Both HL-LHC and DUNE are expected to start data taking by the end of this decade. They pose unprecedented challenges to event data I/O in terms of data rates, event sizes, and event complexity. At the same time, the I/O landscape is becoming more diverse. HPC cluster file systems and object stores, NVMe disk cache layers in analysis facilities, and S3 storage on cloud resources are mixing with traditional XRootD-managed spinning disk pools.The ROOT team will finalize a first production version of the RNTuple binary format by the end of 2024. After this point, ROOT will provide backward compatibility for RNTuple data. This contribution provides an overview of the RNTuple feature set, the related R&D activities and the long-term vision for RNTuple. We report on performance, interface design, tooling, robustness, integration with experiment frameworks, and validation results, as well as recent R&D on parallel reading and writing and exploitation of modern hardware and storage systems. We will give an outlook on possible future features after a first production release.Collaboratively, the IT and EP departments at CERN have launched a formal project within the Research and Computing sector to evaluate the novel data format for physics analysis data utilized in LHC experiments and other fields. This part of the project focuses on validating the scalability of the EOS storage backend during the transition from the over 25 years old TTree production format to the newly developed RNTuple format, using both replicated and erasure-coded storage profiles.

Blomer, Jakob [CERN]↗

Fokker-Planck simulations of fast ion ICRF and electron EC heating in a mirror plasma using CQL3D-m

The CQL3D-m continuum bounce-average Fokker-Planck code is adapted for magnetic mirror plasmas [1] and is now routinely used in no-free-parameter classical integrated modeling of mirror devices [2, 3]. In the present effort, we report on two RF methods of plasma heating in mirror machine. The fast ions (FI) are heated by Fast waves at 2nd-4th harmonic, where FIs originate from neutral beam injection at 45 degrees to the magnetic field. The scenario shows an efficient ion heating near the FI bouncing point. The electrons are heated by X-mode launched from the high magnetic field side towards the resonance. Different from the tokamak applications, CQL3D-m provides an evolving self-consistent ambipolar parallel electric field, which determines the shape of the loss cone and hence an accurate confinement time of both ions and electrons. Also, it includes a description of ion and electron sources and sinks (related to charge exchange and impact ionization) which are updated at every time step. CQL3D-m utilizes a fully nonlinear Coulomb collision operator that is important for the significantly non-Maxwellian ion distributions typically established in mirror plasmas.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Laboratory study of the PFRC-2's initial plasma densification stages

Initial plasma densification by odd-parity rotating magnetic fields (RMF o ) applied to the linear magnetized Princeton field-reversed configuration (PFRC-2) device with fill gases at pressures near 1 mTorr proceeds through two phases: a slow one, characterized by a rise time $τ_{slow}$ ~ 100 $μ$s, followed by a fast one, characterized by $τ_{fast}$ ~ 10 $μ$s. The transition from slow to fast occurs at a line-integral-averaged electron density, t n e , near 2$\times$ 10 11 cm –3 , independent of magnetic field. Here, over most of the range of experimental parameters investigated, as the PFRC-2 axial magnetic field strength was increased, RMF o power decreased, gas fill pressure lowered, or lower atomic mass unit (AMU) fill gas used, the duration of the slow phase lengthened from 50 $μ$s to longer than 10 ms after the RMF o power began. The post-fast-phase maximum n e increases with the fill-gas AMU, exceeding 5 × 10 13 cm –3 for Ar. The slow phase is consistent with atomic physics processes and field-parallel sound-speed losses. The fast phase may be explained by improved axial confinement, possibly augmented by radial or axial contraction of the plasma. Another possible explanation, a large increase in electron temperature, is inconsistent with x-ray emission. The n e behavior is discussed in relation to the E to H transition.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

MicroResonators for Compacts Optical Sensors (μRCOS)

As the demand for continuous, in-situ surveillance of health of systems and environments rapidly increases, miniaturized sensors with fast responses are sought. Optical dielectric resonators supporting Whispering Gallery Modes (WGMs) have exceptional properties, like very high-power density, very narrow spectral linewidth, and extremely small mode volume. As the footprint of these sensors is drastically reduced, measurements of microvolumes samples with dramatic reduction in analysis time are possible, enabling large numbers of parallel analyses using microarrays. The high sensitivity and speed of WGM resonators combined with the ability to detect molecules in their native state have great future potential for basic and applied research such as reliable single molecule, trace-gas detection, environmental monitoring, chemical threat sensing, and biodefense Such appealing peculiarities motivated us developing WMG resonators (WGMRs) for Cavity Enhanced Absorption Spectroscopy (CEAS) for chem-bio detection. Specifically, we designed and batch-fabricated microspheres and fiber tapers for resonances excitation. We successfully detected gases (N 2 and CO 2 ) in customized environmental chambers, and bio-organisms (Inf. A and E. Coli) with integrated microfluidic systems. Background calibration and environmental isolation were always accounted for in proving the performances.

36 MATERIALS SCIENCE↗

MicroResonators for Compacts Optical Sensors (μRCOS)

As the demand for continuous, in-situ surveillance of health of systems and environments rapidly increases, miniaturized sensors with fast responses are sought. Optical dielectric resonators supporting Whispering Gallery Modes (WGMs) have exceptional properties, like very high-power density, very narrow spectral linewidth, and extremely small mode volume. As the footprint of these sensors is drastically reduced, measurements of microvolumes samples with dramatic reduction in analysis time are possible, enabling large numbers of parallel analyses using microarrays. The high sensitivity and speed of WGM resonators combined with the ability to detect molecules in their native state have great future potential for basic and applied research such as reliable single molecule, trace-gas detection, environmental monitoring, chemical threat sensing, and biodefense Such appealing peculiarities motivated us developing WMG resonators (WGMRs) for Cavity Enhanced Absorption Spectroscopy (CEAS) for chem-bio detection. Specifically, we designed and batch-fabricated microspheres and fiber tapers for resonances excitation. We successfully detected gases (N 2 and CO 2 ) in customized environmental chambers, and bio-organisms (Inf. A and E. Coli) with integrated microfluidic systems. Background calibration and environmental isolation were always accounted for in proving the performances.

36 MATERIALS SCIENCE↗

Integrated Task And Data Parallel Programming: Language Design

his research investigates the combination of task and data parallel language constructs within a single programming language. There are an number of applications that exhibit properties which would be well served by such an integrated language. Examples include global climate models, aircraft design problems, and multidisciplinary design optimization problems. Our approach incorporates data parallel language constructs into an existing, object oriented, task parallel language. The language will support creation and manipulation of parallel classes and objects of both types (task parallel and data parallel). Ultimately, the language will allow data parallel and task parallel classes to be used either as building blocks or managers of parallel objects of either type, thus allowing the development of single and multi-paradigm parallel applications. 1995 Research Accomplishments In February I presented a paper at Frontiers '95 describing the design of the data parallel language subset. During the spring I wrote and defended my dissertation proposal. Since that time I have developed a runtime model for the language subset. I have begun implementing the model and hand-coding simple examples which demonstrate the language subset. I have identified an astrophysical fluid flow application which will validate the data parallel language subset. 1996 Research Agenda Milestones for the coming year include implementing a significant portion of the data parallel language subset over the Legion system. Using simple hand-coded methods, I plan to demonstrate (1) concurrent task and data parallel objects and (2) task parallel objects managing both task and data parallel objects. My next steps will focus on constructing a compiler and implementing the fluid flow application with the language. Concurrently, I will conduct a search for a real-world application exhibiting both task and data parallelism within the same program m. Additional 1995 Activities During the fall I collaborated with Andrew Grimshaw and Adam Ferrari to write a book chapter which will be included in Parallel Processing in C++ edited by Gregory Wilson. I also finished two courses, Compilers and Advanced Compilers, in 1995. These courses complete my class requirements at the University of Virginia. I have only my dissertation research and defense to complete.

Grimshaw, Andrew S.↗

Integrated Task and Data Parallel Programming

This research investigates the combination of task and data parallel language constructs within a single programming language. There are an number of applications that exhibit properties which would be well served by such an integrated language. Examples include global climate models, aircraft design problems, and multidisciplinary design optimization problems. Our approach incorporates data parallel language constructs into an existing, object oriented, task parallel language. The language will support creation and manipulation of parallel classes and objects of both types (task parallel and data parallel). Ultimately, the language will allow data parallel and task parallel classes to be used either as building blocks or managers of parallel objects of either type, thus allowing the development of single and multi-paradigm parallel applications. 1995 Research Accomplishments In February I presented a paper at Frontiers 1995 describing the design of the data parallel language subset. During the spring I wrote and defended my dissertation proposal. Since that time I have developed a runtime model for the language subset. I have begun implementing the model and hand-coding simple examples which demonstrate the language subset. I have identified an astrophysical fluid flow application which will validate the data parallel language subset. 1996 Research Agenda Milestones for the coming year include implementing a significant portion of the data parallel language subset over the Legion system. Using simple hand-coded methods, I plan to demonstrate (1) concurrent task and data parallel objects and (2) task parallel objects managing both task and data parallel objects. My next steps will focus on constructing a compiler and implementing the fluid flow application with the language. Concurrently, I will conduct a search for a real-world application exhibiting both task and data parallelism within the same program. Additional 1995 Activities During the fall I collaborated with Andrew Grimshaw and Adam Ferrari to write a book chapter which will be included in Parallel Processing in C++ edited by Gregory Wilson. I also finished two courses, Compilers and Advanced Compilers, in 1995. These courses complete my class requirements at the University of Virginia. I have only my dissertation research and defense to complete.

Grimshaw, A. S.↗