Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “multiple iterations”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 217 records · Page 12

Optimizing Management of Persistent Data Structures in High-Performance Analytics

Large-scale data analytics workflows ingest massive input data into various data structures, including graphs and key-value datastores. These data structures undergo multiple transformations and computations and are typically reused in incremental and iterative analytics workflows. Persisting in-memory views of these data structures enables reusing them beyond the scope of a single program run while avoiding repetitive raw data ingestion overheads. Memory-mapped I/O enables persisting in-memory data structures without data serialization and deserialization overheads. However, memory-mapped I/O lacks the key feature of persisting consistent snapshots of these data structures for incremental ingestion and processing. The obstacles to efficient virtual memory snapshots using memory-mapped I/O include background writebacks outside the application’s control, and the significantly high storage footprint of such snapshots. To address these limitations, we present Privateer, a memory and storage management tool that enables storage-efficient virtual memory snapshotting while also optimizing snapshot I/O performance. Here, we integrated Privateer into Metall, a state-of-the-art persistent memory allocator for C++, and the Lightning Memory-Mapped Database (LMDB), a widely-used key-value datastore in data analytics and machine learning. Privateer optimized application performance by 1.22× when storing data structure snapshots to node-local storage, and up to 16.7× when storing snapshots to a parallel file system. Privateer also optimizes storage efficiency of incremental data structure snapshots by up to 11× using data deduplication and compression.

Computer science↗

Numerical solution for the inviscid supersonic flow in the corner formed by two intersecting wedges.

The inviscid, interference corner flow generated by two intersecting wedges immersed in a supersonic stream is obtained by use of a second-order, shock-capturing, finite-difference approach. The governing equations are solved iteratively in conical coordinates to yield the flow structure consisting of multiple shock and slip surfaces. The numerical results for shock wave and slip surface position and structure, pitot pressure traverses, and surface pressure distributions are compared with experimental data obtained over a wide range of Reynolds numbers. The comparisons show the best agreement with the high Reynolds number (greater than 3,000,000) results for which the boundary layer is turbulent.

Kutler, P.↗

Multi-objective/loading optimization for rotating composite flexbeams

With the evolution of advanced composites, the feasibility of designing bearingless rotor systems for high speed, demanding maneuver envelopes, and high aircraft gross weights has become a reality. These systems eliminate the need for hinges and heavily loaded bearings by incorporating a composite flexbeam structure which accommodates flapping, lead-lag, and feathering motions by bending and twisting while reacting full blade centrifugal force. The flight characteristics of a bearingless rotor system are largely dependent on hub design, and the principal element in this type of system is the composite flexbeam. As in any hub design, trade off studies must be performed in order to optimize performance, dynamics (stability), handling qualities, and stresses. However, since the flexbeam structure is the primary component which will determine the balance of these characteristics, its design and fabrication are not straightforward. It was concluded that: pitchcase and snubber damper representations are required in the flexbeam model for proper sizing resulting from dynamic requirements; optimization is necessary for flexbeam design, since it reduces the design iteration time and results in an improved design; and inclusion of multiple flight conditions and their corresponding fatigue allowables is necessary for the optimization procedure.

Hamilton, Brian K.↗

Thermal Control Subsystem Design for the Avionics of a Space Station Payload

A case study of the thermal control subsystem development for a space based payload is presented from the concept stage through preliminary design. This payload, the Space Acceleration Measurement System 2 (SAMS-2), will measure the acceleration environment at select locations within the International Space Station. Its thermal control subsystem must maintain component temperatures within an acceptable range over a 10 year life span, while restricting accessible surfaces to touch temperature limits and insuring fail safe conditions in the event of loss of cooling. In addition to these primary design objectives, system level requirements and constraints are imposed on the payload, many of which are driven by multidisciplinary issues. Blending these issues into the overall system design required concurrent design sessions with the project team, iterative conceptual design layouts, thermal analysis and modeling, and hardware testing. Multiple tradeoff studies were also performed to investigate the many options which surfaced during the development cycle.

Moran, Matthew E.↗

A Practical Guide to Writing a Radiative Transfer Code

Using our decades-long experience in radiative transfer (RT) code development for Earth science, we endeavor to reduce the knowledge gap of bringing RT from theory to code quickly. Despite numerous classic and recent literature, it is still hard to develop anRT code from scratch within a few weeks. It is equally hard to understand, not to mention modify, an existing “monster” RT code, for which the developer is either located remotely or has retired. Following the format of “Numerical Recipes” by Press et al., we collocate in this paper small pieces of necessary theory with corresponding small pieces of RT code. These are arranged in an order that is natural for code development, which is often opposite of the natural order for laying out the theoretical basis. We focus on the transfer of unpolarized monochromatic solar radiation in a plane-parallel atmosphere over a reflecting surface. Both the surface and the atmosphere are homogeneous (uniform) at all directions. The multiple scattering is numerically solved using the deterministic method of Gauss-Seidel iterations. Except for the presented Python-Numba open-source RT code gsit, the paper does not report any new scientific results, but rather serves as an academic demonstration. If development time is an issue or the reader is familiar with basic concepts of RT theory, we recommend proceeding directly to Sec.3 “RT code development.

multiple light scattering↗

Impurity leakage and radiative cooling in the first nitrogen and neon seeding study in the closed DIII-D SAS configuration

A comparative study of nitrogen versus neon has been carried out to analyze the impact of the two radiative species on power dissipation, SOL impurity distribution, divertor and pedestal characteristics. The experimental results show that N remains compressed in the divertor, thereby providing high radiative losses without affecting the pedestal profiles and displacing carbon as dominant radiator. Neon, instead, radiates more upstream than N thus reducing the power flux through the separatrix leading to a reduced ELM frequency and compression in the divertor. A significant amount of neon is measured in the plasma core leading to a steeper density gradient. The different behavior between the two impurities is confirmed by SOLPS-ITER modeling which for the first time at DIII-D includes multiple impurity species and a treatment of full drifts, currents and neutral–neutral collisions. The impurity transport in the SOL is studied in terms of the parallel momentum balance showing that N is mostly retained in the divertor whereas Ne leaks out consistent with its higher ionization potential and longer mean free path. This is also in agreement with the enrichment factor calculations which indicate lower divertor enrichment for neon. The strong ionization source characterizing the SAS divertor causes a reversal of the main ions and impurity flows. The flow reversal together with plasma drifts and the effect of the thermal force contribute significantly in the shift of the impurity stagnation point affecting impurity leakage. This work provides a demonstration of the impurity leakage mechanism in a closed divertor structure and the consequent impact on pedestal. Since carbon is an intrinsic radiator at DIII-D, in this paper we have also demonstrated the different role of carbon in the N vs Ne seeded cases both in the experiments and in the numerical modeling. Here, carbon contributes more when neon seeding is injected compared to when nitrogen is used. Finally, the results highlight the importance of accompanying experimental studies with numerical modeling of plasma flows, drifts and ionization profile to determine the details of the SOL impurity transport as the latter may vary with changes in divertor regime and geometry. In the cases presented here, plasma drifts and flow reversal caused by high level of closure in the slot upper divertor at DIII-D play an important role in the underlined mechanism.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Application of an advanced computerized structural design system to an arrow-wing supersonic cruise aircraft

A structural design study of an arrow-wing supersonic cruise aircraft has been made using the integrated design system, ATLAS, and a relatively large analytical finite-element model containing 8500 degrees of freedom. This paper focuses on structural design methods developed and used in support of the study with emphasis on aeroelasticity. The use of ATLAS permitted (1) automatic resizing of the wing structure for multiple load conditions, (2) rapid evaluation of aeroelastic effects, and (3) an iterative approach to the correction of flutter deficiencies. The significant results of the study are discussed along with the advantages derived from the use of an advanced structural design system in preliminary design studies.

Robinson, J. C.↗

An analytic radiative transfer model for a coupled atmosphere and leaf canopy

A new analytical radiative transfer model of a leaf canopy is developed that approximates multiple-scattering radiance by a four-stream formulation. The canopy model is coupled to a homogeneous atmospheric model as well as a non-Lambertian lower boundary soil surface. The same four-stream formulation is also used for the calculation of multiple scattering in the atmosphere. Comparisons of radiance derived from the four-stream model with those calculated by an iterative numerical solution of the radiative transfer equation show that the analytic model has a very high accuracy, even with a turbid atmosphere and a very dense canopy in which multiple scattering dominates. Because the coupling of canopy and atmospheric models fully accommodates anisotropic surface reflectance and atmospheric scattering and its effect on directional radiance, the model is especially useful for application to directional radiance and measurements obtained by remote sensing. Retrieval of biophysical parameters using this model is under investigation.

Liang, Shunlin↗

Post-Stall Aerodynamic Modeling and Gain-Scheduled Control Design

A multidisciplinary research e.ort that combines aerodynamic modeling and gain-scheduled control design for aircraft flight at post-stall conditions is described. The aerodynamic modeling uses a decambering approach for rapid prediction of post-stall aerodynamic characteristics of multiple-wing con.gurations using known section data. The approach is successful in bringing to light multiple solutions at post-stall angles of attack right during the iteration process. The predictions agree fairly well with experimental results from wind tunnel tests. The control research was focused on actuator saturation and .ight transition between low and high angles of attack regions for near- and post-stall aircraft using advanced LPV control techniques. The new control approaches maintain adequate control capability to handle high angle of attack aircraft control with stability and performance guarantee.

Wu, Fen↗

Performance Optimization Methods for a Memory-Bound, Unstructured-Grid CFD Application on Massively Parallel GPU Platforms

Computational performance of the FUN3D unstructured-grid computational fluid dynamics (CFD) application on massively parallel GPU environments is memory-bound and highly dependent upon efficient reads from and atomic updates to the irregular cell-, edge-, and node-based data structures. In this talk, we present recent efforts into optimizing select performance-critical kernels on NVIDIA Tesla V100 and A100 GPUs and AMD CDNA MI100 GPUs. A novel use of L2 cache residency controls and asynchronous loads into on-chip shared memory are explored on the A100 GPU for the sparse iterative solver, which is dominated by mixed-precision, sparse matrix vector multiplication. Demonstrations show that these methods improve global memory bandwidth utilization by 13.5% on the A100 GPU. Several techniques are also presented that use registers and/or shared memory to facilitate array transposition and aggregation which combine to reduce the frequency and increase the cache efficiency of floating-point atomic updates to the irregular data structures. These methods are demonstrated to improve the kernel throughput by nearly 500% on select kernels on the AMD MI100 over atomic updates directly to global memory. Overall, both V100 and A100 GPUs outperformed the MI100 GPU on kernels dominated by double-precision atomic updates; however, the techniques demonstrated here reduced the performance gap and improved the MI100 performance.

GPU CPU unstructured CFD memory↗

A quadrilateral vortex method applied to configurations with high circulation

A quadrilateral vortex-lattice method is briefly described for calculating the potential flow aerodynamic characteristics of high-lift configurations. It incorporates an iterative scheme for calculating the deformation of forcefree wakes, including wakes from side edges. The method is applicable to multiple lifting surfaces with part-span flaps deflected, and can include ground effect and wind-tunnel interference. Numerical results, presented for a number of high-lift configurations, demonstrate rapid convergence of the iterative technique. The results are in good agreement with available experimental data.

Maskew, B.↗

Trade-Space Analysis Tool for Constellations (TAT-C)

Traditionally, space missions have relied on relatively large and monolithic satellites, but in the past few years, under a changing technological and economic environment, including instrument and spacecraft miniaturization, scalable launchers, secondary launches as well as hosted payloads, there is growing interest in implementing future NASA missions as Distributed Spacecraft Missions (DSM). The objective of our project is to provide a framework that facilitates DSM Pre-Phase A investigations and optimizes DSM designs with respect to a-priori Science goals. In this first version of our Trade-space Analysis Tool for Constellations (TAT-C), we are investigating questions such as: How many spacecraft should be included in the constellation? Which design has the best costrisk value? The main goals of TAT-C are to: Handle multiple spacecraft sharing a mission objective, from SmallSats up through flagships, Explore the variables trade space for pre-defined science, cost and risk goals, and pre-defined metrics Optimize cost and performance across multiple instruments and platforms vs. one at a time.This paper describes the overall architecture of TAT-C including: a User Interface (UI) interacting with multiple users - scientists, missions designers or program managers; an Executive Driver gathering requirements from UI, then formulating Trade-space Search Requests for the Trade-space Search Iterator first with inputs from the Knowledge Base, then, in collaboration with the Orbit Coverage, Reduction Metrics, and Cost Risk modules, generating multiple potential architectures and their associated characteristics. TAT-C leverages the use of the Goddard Mission Analysis Tool (GMAT) to compute coverage and ancillary data, streamlining the computations by modeling orbits in a way that balances accuracy and performance.TAT-C current version includes uniform Walker constellations as well as Ad-Hoc constellations, and its cost model represents an aggregate model consisting of Cost Estimating Relationships (CERs) from widely accepted models. The Knowledge Base supports both analysis and exploration, and the current GUI prototype automatically generates graphics representing metrics such as average revisit time or coverage as a function of cost.

Science Data Processing↗

Mapping nanocrystal orientations via scanning Laue diffraction microscopy for multi-peak Bragg coherent diffraction imaging

The recent commissioning of a movable monochromator at the 34-ID-C endstation of the Advanced Photon Source has vastly simplified the collection of Bragg coherent diffraction imaging (BCDI) data from multiple Bragg peaks of sub-micrometre scale samples. Laue patterns arising from the scattering of a polychromatic beam by arbitrarily oriented nanocrystals permit their crystal orientations to be computed, which are then used for locating and collecting several non-co-linear Bragg reflections. The volumetric six-component strain tensor is then constructed by combining the projected displacement fields that are imaged using each of the measured reflections via iterative phase retrieval algorithms. Complications arise when the sample is heterogeneous in composition and/or when multiple grains of a given lattice structure are simultaneously illuminated by the polychromatic beam. Here, a workflow is established for orienting and mapping nanocrystals on a substrate of a different material using scanning Laue diffraction microscopy. The capabilities of the developed algorithms and procedures with both synthetic and experimental data are demonstrated. The robustness is verified by comparing experimental texture maps obtained with Laue diffraction microscopy at the beamline with maps obtained from electron back-scattering diffraction measurements on the same patch of gold nanocrystals. Such tools provide reliable indexing for both isolated and densely distributed nanocrystals, which are challenging to image in three dimensions with other techniques.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Decoupled dynamic analysis of combined systems by iterative determination of interface accelerations

A dynamic analysis technique is presented that can be used to determine the response of a discrete model of a large linear structural system composed of multiple substructures. The technique circumvents the costly computation of the modal characteristics of the combined system. This is accomplished by relying on a predictor-corrector scheme to converge iteratively to the interface accelerations of the combined system, while the equations of motions of the individual structures are integrated separately. In this regard, the temporal slopes of the interface accelerations (jerks) are computed at each time point of integration to predict the interface accelerations at the next time point. The proposed technique is exemplified by conducting a Space Shuttle landing loads analysis; the obtained numerical data demonstrate its reliability and efficiency.

Spanos, P. D.↗

Iterative hybrid manufacture of a titanium alloy component

Here, this paper describes an iterative hybrid (additive + subtractive) manufacturing approach for a titanium alloy (Ti6Al4V) part using a laser hotwire directed energy deposition system (LHWDED) and a traditional four-axis milling machine tool. The term iterative hybrid manufacturing is used to described hybrid manufacturing where the additive and subtractive operations occur in multiple stages rather than sequentially. It is currently common to produce an entire part by sequential hybrid manufacturing by additively manufacturing (AM) an entire preform geometry that then requires post processing by another machine tool to create final part features. By contrast, a part produced by iterative hybrid manufacturing (IHM) does not produce the entire preform geometry in a single AM process. Instead, a portion of the entire preform geometry is manufactured by an AM process, then that portion is transferred to another machine tool which creates features in that portion, and then that machined portion is transferred back to the AM machine to complete another AM process. There is no limit to the number of iterations that an IHM process can have. IHM offers several advantages over sequential hybrid manufacturing such as the use of shorter and stiffer subtractive tooling, better access to part geometries that require subtractive processes, and the separation of the AM heat source from the subtractive machine tool. A titanium alloy demonstration part was successfully manufactured by IHM with three iterations using a shared pallet system between the AM machine tool and the subtractive machine tool.

Hybrid manufacturing↗

Computational Issues in Damping Identification for Large Scale Problems

Two damping identification methods are tested for efficiency in large-scale applications. One is an iterative routine, and the other a least squares method. Numerical simulations have been performed on multiple degree-of-freedom models to test the effectiveness of the algorithm and the usefulness of parallel computation for the problems. High Performance Fortran is used to parallelize the algorithm. Tests were performed using the IBM-SP2 at NASA Ames Research Center. The least squares method tested incurs high communication costs, which reduces the benefit of high performance computing. This method's memory requirement grows at a very rapid rate meaning that larger problems can quickly exceed available computer memory. The iterative method's memory requirement grows at a much slower pace and is able to handle problems with 500+ degrees of freedom on a single processor. This method benefits from parallelization, and significant speedup can he seen for problems of 100+ degrees-of-freedom.

Pilkey, Deborah L.↗

Meta-Learning Enhanced Physics-Informed Graph Attention Convolutional Network for Distribution Power System State Estimation

Promptly perceiving distribution system states is challenged by frequent topology changes and uncertain power injections. To address these issues, a Meta-learning enhanced physics-informed graph attention convolutional network (Meta-PIGACN) model is proposed to handle topological variability in distribution system state estimation (DSSE). Specifically, physics information is integrated into the graph convolutional network, enabling a physics-informed edge-weighting process that incorporates physical information to control the aggregation of neighboring nodes. Besides, the graph attention mechanism automatically adjusts the importance of different neighboring nodes, allowing the capture and preservation of inherent system features across varying topologies, thereby improving state estimation accuracy. Furthermore, meta-learning is proposed to acquire empirical knowledge across multiple topologies so that the model can rapidly adapt to new configurations through iterative gradient descent updates even in large-scale systems. In conclusion, the simulation results based on the 33/118/1746-node distribution systems show the high accuracy and efficiency of the proposed model.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Topological solution of bilateral switching networks

Topological method uses the eye as pattern detector to trace path of transmission on truth table. Pathway selection is continually supervised by logician, allowing him to seek planar iterative solution desirable for fabrication of monolithic circuits. Method applies to parity generators, multiple output functions, full adders, and bit comparators.

Mazer, L.↗