Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Compiler frameworks”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

MLIR loop optimizations for High-Level Synthesis: a case study

High-Level Synthesis (HLS) tools simplify the design of hardware accelerators by automatically generating Verilog/VHDL code starting from a general purpose software programming language. They include a wide range of optimization techniques in the process, most of them performed on a low-level intermediate representation (IR) of the code. Introducing optimizations on a higher level of abstraction could significantly contribute to the automated design process results; for example, polyhedral techniques for the manipulation of loops could have a significant impact on the generated accelerators when applied on a specialized IR. We use loop pipelining as a case study to explore the introduction of compiler-based transformations on top of an existing HLS process. We leverage the Multi-Level Intermediate Representation (MLIR) framework and an external scheduler to implement the required transformations, and couple them with existing HLS tools to evaluate the improvements that loop pipelining brings to the performance of generated accelerators. The proposed approach can be integrated with other high-level transformations on the MLIR representation, combining different techniques to obtain pre-optimized inputs for HLS that do not have to rely on a specific backend tool.

Curzel, Serena↗

Caffeine v0.1.0

Caffeine is the CoArray Fortran Framework of Efficient Interfaces to Network Environments. Caffeine aims to produce a parallel runtime library that will support Fortran compilers with a programming-model-agnostic application binary interface (ABI) to various lower-level communication libraries. The current version of Caffeine uses the GASNet-EX networking middleware, also developed at Berkeley Lab. On many combinations of applications and platforms, GASNet-EX outperforms the widely used Message Passing Interface (MPI). Through GASNet-EX's support for communicating between graphics processing units (GPU), GASNet-EX has features that specifically target the emerging, leading-edge exascale computing platforms.

Rouson, Damian↗

Use of Semantic Technology to Create Curated Data Albums

One of the continuing challenges in any Earth science investigation is the discovery and access of useful science content from the increasingly large volumes of Earth science data and related information available online. Current Earth science data systems are designed with the assumption that researchers access data primarily by instrument or geophysical parameter. Those who know exactly the data sets they need can obtain the specific files using these systems. However, in cases where researchers are interested in studying an event of research interest, they must manually assemble a variety of relevant data sets by searching the different distributed data systems. Consequently, there is a need to design and build specialized search and discover tools in Earth science that can filter through large volumes of distributed online data and information and only aggregate the relevant resources needed to support climatology and case studies. This paper presents a specialized search and discovery tool that automatically creates curated Data Albums. The tool was designed to enable key elements of the search process such as dynamic interaction and sense-making. The tool supports dynamic interaction via different modes of interactivity and visual presentation of information. The compilation of information and data into a Data Album is analogous to a shoebox within the sense-making framework. This tool automates most of the tedious information/data gathering tasks for researchers. Data curation by the tool is achieved via an ontology-based, relevancy ranking algorithm that filters out nonrelevant information and data. The curation enables better search results as compared to the simple keyword searches provided by existing data systems in Earth science.

Ramachandran, Rahul↗

Use of Semantic Technology to Create Curated Data Albums

One of the continuing challenges in any Earth science investigation is the discovery and access of useful science content from the increasingly large volumes of Earth science data and related information available online. Current Earth science data systems are designed with the assumption that researchers access data primarily by instrument or geophysical parameter. Those who know exactly the data sets they need can obtain the specific files using these systems. However, in cases where researchers are interested in studying an event of research interest, they must manually assemble a variety of relevant data sets by searching the different distributed data systems. Consequently, there is a need to design and build specialized search and discovery tools in Earth science that can filter through large volumes of distributed online data and information and only aggregate the relevant resources needed to support climatology and case studies. This paper presents a specialized search and discovery tool that automatically creates curated Data Albums. The tool was designed to enable key elements of the search process such as dynamic interaction and sense-making. The tool supports dynamic interaction via different modes of interactivity and visual presentation of information. The compilation of information and data into a Data Album is analogous to a shoebox within the sense-making framework. This tool automates most of the tedious information/data gathering tasks for researchers. Data curation by the tool is achieved via an ontology-based, relevancy ranking algorithm that filters out non-relevant information and data. The curation enables better search results as compared to the simple keyword searches provided by existing data systems in Earth science.

Ramachandran, Rahul↗

A Multi-Architecture Approach for Implicit Computational Fluid Dynamics on Unstructured Grids

High-performance computing (HPC) architectures are trending toward manycore paradigms such as graphics processing units (GPUs). Approximately half of the top 100 publicly disclosed supercomputers in the world utilize GPU accelerators for performance. This is in contrast to a decade ago, where there were only a few such machines in the top 100. It is not currently possible to compile and run legacy central processing unit (CPU) software efficiently on GPUs without significant refactoring. Though a number of frameworks offering performance portability exist, none offer a standardized specification that is supported by all major hardware vendors. Additionally, experiences show that obtaining a high percentage of peak performance often requires architecture-specific code. This work details a pragmatic multi-architecture computational fluid dynamics library focused on aerospace problems across the speed range from low subsonic to hypersonic flows involving thermochemical nonequilibrium. A thin abstraction layer above NVIDIA CUDA C++ is utilized, which enables primarily single-source software currently capable of running efficiently on multicore CPUs, NVIDIA GPUs, AMD GPUs, and Intel GPUs. Results on various problems of interest across the speed range are presented and performance is compared between various architectures.

GPU↗

A Multi-Architecture Approach for Implicit Computational Fluid Dynamics on Unstructured Grids

High-performance computing (HPC) architectures are trending toward manycore paradigms such as graphics processing units (GPUs). Approximately half of the top 100 publicly disclosed supercomputers in the world utilize GPU accelerators for performance. This is in contrast to a decade ago, where there were only a few such machines in the top 100. It is not currently possible to compile and run legacy central processing unit (CPU) software efficiently on GPUs without significant refactoring. Though a number of frameworks offering performance portability exist, none offer a standardized specification that is supported by all major hardware vendors. Additionally, experiences show that obtaining a high percentage of peak performance often requires architecture-specific code. This work details a pragmatic multi-architecture computational fluid dynamics library focused on aerospace problems across the speed range from low subsonic to hypersonic flows involving thermochemical nonequilibrium. A thin abstraction layer above NVIDIA CUDA C++ is utilized, which enables primarily single-source software currently capable of running efficiently on multicore CPUs, NVIDIA GPUs, AMD GPUs, and Intel GPUs. Results on various problems of interest across the speed range are presented and performance is compared between various architectures.

GPU↗

Meta-analysis of biogas upgrading to renewable natural gas through biological CO 2 conversion

Biogas upgrading through CO 2 conversion by hydrogenotrophic methanogenesis is receiving an increasing attention worldwide because of the demand for renewable natural gas. Herein, a holistic and statistical study of the operation conditions, driving forces, performances, and potential implementation of biogas upgrading via biological CO 2 conversion was conducted. Based on a systematic review and meta-analysis of 46 existing publications that were selected from 1475 papers, we have compiled a global dataset of CO 2 bioconversion biogas upgrading, encompassing 308 study cases. Subsequently, we employed a rigorous analytical framework incorporating data processing and mixed effects linear regression analysis to examine the dataset. This analysis revealed a significant positive relationship between the H 2 :CO 2 ratio and the methane percentage in the upgraded biogas when using the study as a random effect. Furthermore, we performed meta-analysis on observations taken when the ratio was close to 4:1 and found that ex situ reactors (91.93% [88.11%, 95.75%]) can perform better than in situ reactors (84.74% [80.69%, 88.80%]). No evidence of differential performance was found based on the present dataset between different temperature regimes or operation modes. Furthermore, those findings establish a database that will contribute to a deeper understanding of the biogas upgrading via biological CO 2 conversion.

Biogas upgrading↗

Boosting RDataFrame performance with transparent bulk event processing

RDataFrame is ROOT’s high-level interface for Python and C++ data analysis. Since it first became available, RDataFrame adoption has grown steadily and it is now poised to be a major component of analysis software pipelines for LHC Run 3 and beyond. Thanks to its design inspired by declarative programming principles, RDataFrame enables the development of highperformance, highly parallel analyses without requiring expert knowledge of multi-threading and I/O: user logic is expressed in terms of self-contained, small computation kernels tied together by a high-level API. This design completely decouples analysis logic from its actual execution, and opens several interesting avenues for workflow optimization. In particular, in this work we explore the benefits of moving internal data processing from an event-by-event to a bulkby-bulk loop. This refactoring dramatically reduces the framework’s runtime overheads; in collaboration with the I/O layer it improves data access patterns; it exposes information that optimizing compilers might use to auto-vectorize the invocation of user-defined computations; finally, while existing user-facing interfaces remain unaffected, it becomes possible to additionally offer interfaces that explicitly expose bulks of events, useful e.g. for the injection of GPU kernels into the analysis workflow. In order to inform similar future R&D, design challenges will be presented, as well as an investigation of the relevant timememory trade-off backed by novel performance benchmarks.

Guiraud, Enrico↗

A Systematic Review and Meta-analysis of the Potential Non-human Animal Reservoirs and Arthropod Vectors of the Mayaro Virus

Improving our understanding of Mayaro virus (MAYV) ecology is critical to guide surveillance and risk assessment. We conducted a PRISMA-adherent systematic review of the published and grey literature to identify potential arthropod vectors and non-human animal reservoirs of MAYV. We searched PubMed/MEDLINE, Embase, Web of Science, SciELO and grey-literature sources including PAHO databases and dissertation repositories. Studies were included if they assessed MAYV virological/immunological measured occurrence in field-caught, domestic, or sentinel animals or in field-caught arthropods. We conducted an animal seroprevalence meta-analysis using a random effects model. We compiled granular georeferenced maps of non-human MAYV occurrence and graded the quality of the studies using a customized framework. Overall, 57 studies were eligible out of 1523 screened, published between the years 1961 and 2020. Seventeen studies reported MAYV positivity in wild mammals, birds, or reptiles and five studies reported MAYV positivity in domestic animals. MAYV positivity was reported in 12 orders of wild-caught vertebrates, most frequently in the orders Charadriiformes and Primate. Sixteen studies detected MAYV in wild-caught mosquito genera including Haemagogus, Aedes, Culex, Psorophora, Coquillettidia, and Sabethes. Vertebrate animals or arthropods with MAYV were detected in Brazil, Panama, Peru, French Guiana, Colombia, Trinidad, Venezuela, Argentina, and Paraguay. Among non-human vertebrates, the Primate order had the highest pooled seroprevalence

Mayaro virus↗

Invited: Bambu: an Open-Source Research Framework for the High-Level Synthesis of Complex Applications

This paper presents the open-source High-Level Synthesis research framework Bambu. The framework provides an open-source starting point to experiment with new ideas across High-Level Synthesis, high-level verification and debugging, FPGA/ASIC design, design flow space exploration, and parallel hardware accelerator design. The tool accepts as input standard C/C++ specifications and compiler intermediate representations (IRs) coming from the well-known Clang/LLVM and GCC com- pilers. The broad spectrum and flexibility of input formats allow the electronic design automation (EDA) research community to explore and integrate new transformations and optimizations. The easily extendable modular framework already includes many op- timizations and HLS benchmarks. The integration with synthesis and verification backends (commercial and open-source) allows researchers to quickly test any new finding and easily obtain performance and resource usage metrics for a given application. Different FPGA devices are supported from several different vendors: AMD/XILINX, Intel/Altera, Lattice Semiconductor, and NanoXplore. Finally, integration with the OpenRoad open-source end-to-end silicon compiler perfectly fits with the recent push towards open-source EDA.

Ferrandi, Fabrizio↗

Runtime Verification with Ogma

Ultra-critical systems require high-level assurance, which cannot always be guaranteed in compile time. The use of runtime verification (RV) enable monitoring these systems in runtime, to detect property violations early and limit their potential consequences. However, the introduction of monitors in ultra-critical systems poses a challenge, as failures and delays in the RV subsystem could affect other subsystems and threaten the mission as a whole. In this talk we discuss two systems: NASA's Ogma, a tool to transform high-level specifications into monitoring code, and Copilot, a runtime verification framework for real-time embedded systems. The toolchain can be used to translate structured natural language requirements into C code with static memory requirements, which can be compiled to run on embedded hardware.

Ogma↗

Runtime Verification with Ogma

Ultra-critical systems require high-level assurance, which cannot always be guaranteed in compile time. The use of runtime verification (RV) enable monitoring these systems in runtime, to detect property violations early and limit their potential consequences. However, the introduction of monitors in ultra-critical systems poses a challenge, as failures and delays in the RV subsystem could affect other subsystems and threaten the mission as a whole. In this talk we discuss two systems: NASA's Ogma, a tool to transform high-level specifications into monitoring code, and Copilot, a runtime verification framework for real-time embedded systems. The toolchain can be used to translate structured natural language requirements into C code with static memory requirements, which can be compiled to run on embedded hardware.

Ogma↗

U.S. Industry Opportunities for Advanced Nuclear Technology Development (Phase III)

This is the third phase of a work scope which has been focused on providing a framework for and preserving key experimental programs and experiences which are critical to the licensing basis of currently operating nuclear reactors or which could be used in the safety basis for the next generation of reactors. The first phase of this effort, documented in Reference 1, focused on compiling a list of key experimental programs through an international survey of reactor safety professionals working in licensing, design, and academia. The second phase in this effort, documented in Reference 2, focused on creating a searchable database framework in which to organize the results from the international survey and to perform detailed research on several key programs to provide the framework for how to categorize the references which could be located. The purpose of the third phase is to perform a high-level research effort on each experiment/experience and to determine if sufficient data, reports, and results have already been captured to consider the program archived for future generators of nuclear professionals.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Risk-based area of review estimation in overpressured reservoirs to support injection well storage facility permit requirements for CO 2 storage projects

This paper by the Energy & Environmental Research Center presents a workflow and modeling approach for delineating a risk-based area of review (AOR) to support a U.S. Environmental Protection Agency (EPA) Class VI permit for a carbon dioxide (CO 2 ) storage project. The approach combines semianalytical solutions for estimating formation fluid leakage through a hypothetical leaky wellbore with the results of numerical reservoir simulations to define the AOR. The modeling utilizes 1) semianalytical solutions from the peer-reviewed literature for formation fluid leakage through abandoned wellbores by Raven (1990) and Avci (1994), 2) a FORTRAN model compiled and described in Cihan et al. (2011, 2012) called ASLMA (Analytical Solution for Leakage in Multilayered Aquifers), and 3) a computational framework for estimating a risk-based AOR first proposed by Oldenburg et al. (2014, 2016). Therefore, the approach builds upon well-established research and underlying hydrogeological principles that have been upheld for nearly three decades. Moreover, the ASLMA model has been broadly applied to an array of storage projects. The work presented herein extends these earlier works using a custom wrapper written in the software environment, R (R Core Team, 2020), which was developed to perform multiple runs of the ASLMA model using given ranges for one or more input parameters. In addition, the current work simulates the pressure buildup within the storage reservoir in response to CO 2 injection using a compositional simulator to better accommodate the temporospatial evolution of pressure buildup within the storage reservoir that is more accurately modeled using a heterogeneous geologic model and a compositional simulator that accounts for the multiphase interactions. The workflow is demonstrated using a case study for a 180,000-metric-ton-per-year storage project located in the PCOR (Plains CO 2 Reduction) Partnership region. For the storage project evaluated here, under the scenario where the leaky wellbore is open to a saline aquifer (thief zone) between the overlying seal (cap rock) and the underground sources of drinking water (USDW), the risk-based AOR essentially collapses to the areal extent of the CO 2 plume in the storage reservoir because the pressure buildup in the storage reservoir beyond the CO 2 plume is insufficient to drive formation fluids up a hypothetical leaky wellbore into the USDW. However, even under the conservative assumption that the leaky wellbore is not open to a thief zone, beyond the areal extent of the CO 2 plume, the incremental leakage is less than 400 m 3 over 20 years, which represents ~0.0001% or less of the total volume of water contained within the USDW rock volume. As discussed in the text, the threshold criterion for defining the risk-based AOR is site-specific and should be informed by the results of the sensitivity analysis and available site characterization data. The approach outlined in this paper is designed to be protective of USDWs and, therefore, comply with the Safe Drinking Water Act requirements and provisions for the U.S. EPA Class VI Underground Injection Control (UIC) Program (Class VI Rule) and North Dakota Administrative Code Chapter 43-05-01.

54 ENVIRONMENTAL SCIENCES↗

Central Asia Seismic Hazard Assessment (CASHA): A Probabilistic Seismic Hazard Assessment for Kazakhstan, Kyrgyzstan and Tajikistan

Probabilistic seismic hazard assessments (PSHA) underpin the calculation of earthquake loads in most building codes around the world. In Central Asia, the building codes are slowly being updated to incorporate some of the contemporary concepts of seismic hazard representation. There is also a regional desire to coordinate hazard assessments and building code modernization. However, some challenges remain. Expertise in the region related to seismic hazard assessments is still largely compartmentalised, requiring a significant amount of training and capacity building in seismic hazard assessment related topics. In addition, there are vast amounts of seismic data (bulletin and waveforms), both from analogue and digital eras, that the region’s countries stored but until recently did not use or share among themselves or with the broader seismological community around the world. Finally, after the collapse of the Soviet Union in the 1990s, many of the countries’ seismic networks suffered a major setback with the lack of attention and budget to update existing equipment and installation of new instruments. In order to address these issues, the United States Department of Energy through Lawrence Livermore National Laboratory (LLNL) initiated a project in 2016 to engage and train local scientists in Central Asia to install new equipment, to enhance the quality of seismic monitoring and reporting, to improve and harmonise the regional earthquake catalogue, and to conduct national probabilistic seismic hazard assessments using the new and improved datasets. To achieve the seismic hazard assessment related goals, a series of workshops were held in Almaty, Kazakhstan; Bishkek, Kyrgyzstan; and Dushanbe, Tajikistan from 2016 until 2020. During the time that the COVID-19 pandemic restricted travel, workshops continued online (22 online workshops were hosted in two years). Finally, in May 2022, an in-person workshop in Istanbul, Turkey brought together all project participants along with civil engineers engaged with building code activities in their respective countries, providing a platform to discuss the implementation of the hazard models into updates of building codes in each country, as well as to discuss model parameters, sensitivity analyses and model results in terms of hazard maps, uniform hazard spectra and hazard deaggregations. The workshops were a combination of lectures and hands-on exercises, and included international participation as well as local scientists and engineers. The workshops served several purposes, including training, coordination of data collection, interactions between local earth scientists and engineers, and brainstorming and knowledge exchange among local and international experts. This report outlines the new earthquake catalogue compilation effort and the PSHA project undertaken in Kyrgyzstan, Tajikistan, and Kazakhstan as part of this initiative. The southern part of this region is tectonically active with moderate to high levels of both shallow crustal seismic activity and occurrence of deeper earthquakes under the Hindu Kush and Pamir mountain ranges. Deeper earthquakes also occur near southwestern Kazakhstan, under the eastern Greater Caucasus and Caspian Sea. Large portions of central and northern Kazakhstan, on the other hand, are in stable continental regions with low levels of seismic activity. This study systematically compiles and improves all available data on local seismicity, active faults, and ground motion attenuation characteristics of the region; and builds a framework to enable a contemporary PSHA to be carried out with the engagement of local scientists. While the project was regional, the seismic hazard assessments are primarily driven by the countries’ own national preferences and understanding of data collection, interpretation, and validation of results.

58 GEOSCIENCES↗

Tool for Generation of MAC/GMC Representative Unit Cell for CMC/PMC Analysis

This document describes a recently developed analysis tool that enhances the resident capabilities of the Micromechanics Analysis Code with the Generalized Method of Cells (MAC/GMC) 4.0. This tool is especially useful in analyzing ceramic matrix composites (CMCs), where higher fidelity with improved accuracy of local response is needed. The tool, however, can be used for analyzing polymer matrix composites (PMCs) as well. MAC/GMC 4.0 is a composite material and laminate analysis software developed at NASA Glenn Research Center. The software package has been built around the concept of the generalized method of cells (GMC). The computer code is developed with a user friendly framework, along with a library of local inelastic, damage, and failure models. Further, application of simulated thermomechanical loading, generation of output results, and selection of architectures to represent the composite material have been automated to increase the user friendliness, as well as to make it more robust in terms of input preparation and code execution. Finally, classical lamination theory has been implemented within the software, wherein GMC is used to model the composite material response of each ply. Thus, the full range of GMC composite material capabilities is available for analysis of arbitrary laminate configurations as well. The primary focus of the current effort is to provide a graphical user interface (GUI) capability that generates a number of different user-defined repeating unit cells (RUCs). In addition, the code has provisions for generation of a MAC/GMC-compatible input text file that can be merged with any MAC/GMC input file tailored to analyze composite materials. Although the primary intention was to address the three different constituents and phases that are usually present in CMCs-namely, fibers, matrix, and interphase-it can be easily modified to address two-phase polymer matrix composite (PMC) materials where an interphase is absent. Currently, the tool capability includes generation of RUCs for square packing, hexagonal packing, and random fiber packing as well as RUCs based on actual composite micrographs. All these options have the fibers modeled as having a circular cross-sectional area. In addition, a simplified version of RUC is provided where the fibers are treated as having a square cross section and are distributed randomly. This RUC facilitates a speedy analysis using the higher fidelity version of GMC known as HFGMC. The first four mentioned options above support uniform subcell discretization. The last one has variable subcell sizes due to the primary intention of keeping the RUC size to a minimum to gain the speed ups using the higher fidelity version of MAC. The code is implemented within the MATLAB (The Mathworks, Inc., Natick, MA) developmental framework; however, a standalone application that does not need a priori MATLAB installation is also created with the aid of the MATLAB compiler.

Materials Engineering↗

A Pulse Generation Framework with Augmented Program-aware Basis Gates and Criticality Analysis

Near-term intermediate scale quantum (NISQ) de- vices are subject to considerable noise and short coherence time. Consequently, it is critical to minimize circuit execution latency. Traditionally, each basis gate of a transpiled circuit is decoded into a fixed episode of the device control pulses. Recently, people started to investigate merged pulse generation for customized gates through quantum optimal control (QOC). However, existing QOC approaches face the challenges of (i) restricted search space due to prohibitive compilation overhead; (ii) suboptimal end-to-end performance due to aggressive local optimization and falsely introduced dependency among the customized gates; (iii) inadequate adaptivity towards system calibration, which is critical for NISQ devices. In this work, we propose PAQOC, a novel QOC framework that can (i) automatically detect frequently encountered gate patterns in the logical circuit by modeling the problem as a subgraph mining process and reuse these patterns to enable much larger search space exploration (i.e., program aware); (ii) systemically construct customized gate-set based on the impact to the overall program latency (i.e., criticality-aware); and (iii) quickly adapt to system re-calibration thanks to the small-scale pattern-based gate generation (i.e., adaptivity-aware). PAQOC achieves a good tradeoff between circuit performance and compilation time, allowing fully automatic, single stop, ad- hoc customized pulse generation for more efficient execution of user programs on NISQ devices. Evaluations using fifteen applications show that PAQOC can achieve on average 1.95× speedup of the circuit latency and achieve on average 36.7% reduction in compilation overhead. With PAQOC, circuits can run faster with reduced noise, allowing deeper circuits to be tested within the coherence time of present NISQ platforms.

Chen, Yanhao↗

Recommendations for Distributed Energy Resource Access Control

Cybersecurity for internet - connected Distributed Energy Resources (DER) is essential for the safe and reliable operation of the US power system. Many facets of DER cybersecurity are currently being investigated within different standards development organizations, research communities, and industry committees to address this critical need. This report covers DER access control guidance compiled by the Access Controls Subgroup of the SunSpec/Sandia DER Cybersecurity Workgroup. The goal of the group was to create a consensus - based technical framework to minimize the risk of unauthorized access to DER systems. The subgroup set out to define a strict control environment where users are authorized to access DER monitoring and control features through three steps: (a) user is identified using a proof-of-identity, (b) the user is authenticated by a managed database, (c) and the user is authorized for a specific level of access. DER access control also provides accountability and nonrepudiation within the power system control environment that can be used for forensic analysis and attribution in the event of a cyber-attack. This paper covers foundational requirements for a DER access control environment as well as offering a collection of possible policy, model, and mechanism implementation approaches for IEEE 1547-mandated communication protocols.

45 MILITARY TECHNOLOGY, WEAPONRY, AND NATIONAL DEF↗