Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “github”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Code Coverage Reporter

The code-coverage-reporter is a JavaScript GitHub Action that may be called from within GitHub workflows. Its role is to parse a code coverage report file in the LCOV format and translate the contents into a summary that gets added to individual Pull Requests. It is fully reusable and repository-agnostic.

Curley, Jacob [Fermilab]↗

teemi: An open-source literate programming approach for iterative design-build-test-learn cycles in bioengineering

Synthetic biology dictates the data-driven engineering of biocatalysis, cellular functions, and organism behavior. Integral to synthetic biology is the aspiration to efficiently find, access, interoperate, and reuse high-quality data on genotype-phenotype relationships of native and engineered biosystems under FAIR principles, and from this facilitate forward-engineering strategies. However, biology is complex at the regulatory level, and noisy at the operational level, thus necessitating systematic and diligent data handling at all levels of the design, build, and test phases in order to maximize learning in the iterative design-build-test-learn engineering cycle. To enable user-friendly simulation, organization, and guidance for the engineering of biosystems, we have developed an open-source python-based computer-aided design and analysis platform operating under a literate programming user-interface hosted on Github. The platform is called teemi and is fully compliant with FAIR principles. In this study we apply teemi for i) designing and simulating bioengineering, ii) integrating and analyzing multivariate datasets, and iii) machine-learning for predictive engineering of metabolic pathway designs for production of a key precursor to medicinal alkaloids in yeast. The teemi platform is publicly available at PyPi and GitHub.

59 BASIC BIOLOGICAL SCIENCES↗

HERO WEC 2024 Hydraulic Configuration Deployment Data

The following submission includes raw and processed data from the in water deployment of NREL's Hydraulic and Electric Reverse Osmosis Wave Energy Converter (HERO WEC), in the form of parquet files, TDMS files, CSV files, bag files and MATLAB workspaces. This dataset was collected in March 2024 at the Jennette's pier test site in North Carolina. This submission includes the following: - Data description document (HERO WEC FY24 Hydraulic Deployment Data Descriptions.doc) - This document includes detailed descriptions of the type of data and how it was processed and/or calculated. - Processed MATLAB workspace - The processed data is provided in the form of a single MATLAB workspace containing data from the full deployment. This workspace contains data from all sensors down sampled to 10 Hz along with all array Value Added Products (VAPs). - MATLAB visualization scripts - The MATLAB workspaces can be visualized using the file "HERO_WEC_2024_Hydraulic_Config_Data_Viewer.m/mlx". The user simply needs to download the processed MATLAB workspaces, specify the desired start and end times and run this file. Both the .m and .mlx file format has been provided depending on the user's preference. - Summary Data - The fully processed data was used to create a summary data set with averages and important calculations performed on 30-minute intervals to align with the intervals of wave resource data reported from nearby CDIP ocean observing buoys located 20km East of Jennette's pier and 40km Northeast of Jennette's pier. The wave resource data provided in this data set is to be used for reference only due the difference in water depth and proximity to shore between the Jennette's pier test site and the locations of the ocean observing buoys. This data is provided in the Summary Data zip folder, which includes this data set in the form of a MATLAB workspace, parquet file, and excel spreadsheet. - Processed Parquet File - The processed data is provided in the form of a single parquet file containing data from all HERO WEC sensors collected during the full deployment. Data in these files has been down sampled to 10 Hz and all array VAPs are included. - Interim Filtered Data - Raw data from each sensor group partitioned into 30-minute parquet files. These files are outputs from an intermediate stage of data processing and contain the raw data with no Quality Control (QC) or calculations performed in a format that is easier to use than the raw data. - Raw Data - Raw, unprocessed data from this deployment can be found in the Raw Data zip folder. This data is provided in the form of TDMS, CSV, and bag files in the original format output by the MODAQ system. - Python Data Processing Script - This links to an NREL public github repository containing the python script used to go from raw data to fully processed parquet files. Additional documentation on how to use this script is included in the github repository. This data set has been developed by the National Renewable Energy Laboratory, operated by Alliance for Sustainable Energy, LLC, for the U.S. Department of Energy (DOE) under Contract No. DE-AC36-08GO28308. Funding provided by the U.S. Department of Energy Office of Energy Efficiency and Renewable Energy Water Power Technologies Office.

16 TIDAL AND WAVE POWER↗

ESS-DIVE Unoccupied Aerial Systems (UAS) Reporting Format v1

Here we present documentation of the ESS-DIVE reporting format for Unoccupied Aerial System (UAS) data and metadata. This reporting format provides guidance to data contributors on how to store data to maximize their discoverability, facilitate their efficient reuse, and add value to individual datasets. For data users, the reporting format will better allow data repositories to optimize data search and extraction, and more readily integrate similar data into harmonized synthesis products. The reporting format provides templates and guidance for the reporting of metadata for UAS experimental campaigns, individual flights, platform and sensor description. To improve data access and discoverability, the reporting format proposes a data description scheme of Levels based on the degree of processing, where Level 0 includes raw data, through to Level 3 being derived data end products. A range of examples of data types for each Level are given, with suggested file naming schemes. The reporting format presented here is intended to form a foundation for future development that will accommodate new UAS technologies and approaches to data access and use in the future. The reporting format documentation is maintained and updated on the ESS-DIVE Community Space GitHub at https://github.com/ess-dive-community/essdive-uas. This data package is the first published version of this reporting format, and comprises a zip file of the complete content of https://github.com/ess-dive-community/essdive-uas v1.0. The zip contains the reporting format description, instructions and variable definitions in GitHub markdown language (*.md) and metadata templates in csv format. The reporting format is designed to be compatible with other ESS-DIVE formats, and it is specifically recommended that this reporting format be used in conjunction with the File-level metadata (FLMD) and comma separated values (csv) reporting formats for submission to the ESS-DIVE repository.

54 ENVIRONMENTAL SCIENCES↗

useeio-infra-app

This GitHub repository hosts the source code to an RShiny web application that implements the U.S. Environmentally-Extended Input-Output (USEEIO) model in order to conduct a screening-level life cycle analysis. The USEEIO model is an open-source model developed and maintained by the Environmental Protection Agency (EPA), along with colleagues and contractors, and is available on GitHub (https://github.com/USEPA/USEEIO). The USEEIO model incorporates environmental data into pre-existing Economic Input-Output Models, which look at the interdependence of different economic industries. In this web application, the focus is on energy infrastructure applications; however, it can be used to implement USEEIO in any area for which this model is applicable.

AS↗

Development and Porting of Nuclear Reactor Computational Models for the NRIC Virtual Test Bed in FY23

The U.S. DOE’s National Reactor Innovation Center (NRIC) hosts an open-source website and associated GitHub repository called the Virtual Test Bed (VTB) on which computational models for advanced reactors are documented and shared with the reactor community. In FY23 under the NRIC program, computational models for nuclear reactor analysis were developed and contributed to the VTB Github repository during FY23 by contributors at Argonne National Laboratory. Other models developed outside of NRIC were also ported to the VTB as part of this work, serving a broader mission to make computational reactor analysis models more widely available to the reactor community. The model development activities focused on demonstration of gas-cooled microreactor models including computational fluid dynamics simulations of flow through an industry-inspired air jacket design, and simulation of multiphysics transients for a gas-cooled microreactor assembly. Additionally, development of a molten salt reactor (led at Idaho National Laboratory) was supported through Argonne’s expertise in multigroup cross sections generation. Models for these two reactor types were targeted due to their relevancy to NRIC’s Demonstration of Microreactor Experiments (DOME) and Laboratory for Operation and Testing in the U.S. (LOTUS) physical test beds, which are slated to host microreactor and molten salt reactor experiments. The model porting activities consisted of developing detailed documentation for several physics models which originate from the Nuclear Energy Advanced Modeling and Simulation (NEAMS) program. We include only contributions from Argonne National Laboratory in this report.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Using Apptainer in a Pilot-based Distributed Workload

GlideinWMS is a pilot and pressure-based workload manager for distributed scientific computing. Many experiments like CMS and Fermilab’s Neutrino experiments use it to provision elastic clusters for their analysis and simulations, split into close to a million concurrent jobs. Most user jobs require containers, and the pilots use Apptainer to set up the desired platform. For the pilots that run as regular batch jobs, Apptainer is safer, lighter, and easier to use than other containerization solutions. Many images used by the pilots are expanded SIF images distributed via the CernVM-FS: this combination is very efficient. At Fermilab, for example, we store on GitHub Dockerfiles that mimic the platform in the worker nodes of local clusters. GitHub workflows build and push the images to Docker Hub, and a service periodically pulls and converts them to the expanded SIF images in the CernVM-FS, so the scientists can find a familiar environment everywhere. Apptainer has also been used to run services inside the pilot jobs, like benchmarks that characterize the worker node being used, or a Triton Inference Server that allows sharing a GPU with all the jobs that run in parallel on a node.

Mambelli, Marco [Fermilab] (ORCID:0000000294892681↗

cnor_pub: R code and data for nitrous oxide synthesis by purified bacterial cNOR

This zipped archive of a GitHub repository includes the experimental data (csv files) collected for the reduction of NO to N2O by purified Paracoccus denitrificans cytochrome c nitric oxide reductase (cNOR) and the R code (qmd files) used to analyze these data. A link to the corresponding GitHub repository is also provided.

Hegg, Eric L. [GLBRC - Michigan State University]↗

MSD CoP Webinar: Using Meta-Repositories To Facilitate Open Science in MSD Research

Context: This webinar was hosted by the MultiSector Dynamics Community of Practice (MSD CoP; https://multisectordynamics.org). Abstract: In this webinar, we will describe the use of a GitHub meta-repository, for documenting and disseminating the tools and data supporting MSD publications. The goal is to make it easier for others to understand the flow of data and code through your experiment and to be able to reproduce your results and figures with only the information you have provided for them. The webinar will cover the role of open science in the MSD community, the origins and purpose of meta-repositories, and step-by-step instructions and best practices for building a meta-repository starting from the GitHub template (https://github.com/IMMM-SFA/metarepo). We will also discuss how to leverage MSD-LIVE (https://msdlive.org/) in your meta-repository, provide links to numerous examples you can learn from, and discuss the role of meta-repositories in the IM3 project's open science mandates. Presenters : Chris R. Vernon, Casey D. Burleyson, Jennie Rice, and Mengqi Zhao Moderator: Pat M. Reed (MSD CoP Facilitation Team) This webinar was held on: February 22nd, 2024 from 2-3 PM ET

Open Science↗

WaterTAP3 Model Input Data for NAWI's Eight Source Water Baseline Analyses

This folder contains the input data for the WaterTAP3 model that was used for the eight NAWI (National Alliance for Water Innovation) source water baselines studies published in the Environmental Science and Technology special issue: Technology Baselines and Innovation Priorities for Water Treatment and Supply. There are also eight other separate DAMS submissions, one per source water, that include the model results for the published studies. In this data submission, all model inputs across the eight baselines are included. The data structure and content are described in a README.txt file. For more details on how to use the data in WaterTAP3 please refer to the model documentation and GitHub site found at "WaterTAP3 Github" linked in the submission resources.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

FORCE Regression Testing

Via programs including the Light Water Reactor Sustainability and Integrated Energy Systems, the U.S. Department of Energy has invested in the Framework for Optimization of ResourCes and Economics (FORCE) software framework (Idaho National Laboratory 2024a) for the technical and economic analysis of nuclear-integrated energy systems (IES). Nuclear IES expand the use of nuclear from traditional baseload electricity generation to a flexible and adaptive source of combined heat and power. Nuclear heat can be used in the production of a variety of energy currencies such as hydrogen and ammonia as well as other heat applications including water desalination and district heating. FORCE is designed with the intent to provide interconnected analysis tools that enable the accurate technical and economic assessment of specific nuclear IES configurations for individual energy markets. FORCE consists of three main analysis pathways: HYBRID (Idaho National Laboratory 2024b), which contains high-resolution physical models for IES; Holistic Energy Resource Optimization Network (HERON) (Idaho National Laboratory 2024c), which analyzes IES long-term economic viability; and Optimization of Real-time Capacity Allocation (ORCA) (Idaho National Laboratory 2024d), designed for real-time control of IES via digital twins and optimal decision making, including autonomous and remote operation research. Development of the FORCE ecosystem is guided by three pillars: capability, which assures that the computational requirements of IES analysis are met by the software tools; reliability, which provides for consistent code performance and expected behaviors; and accessibility, which lowers the barrier to entry for using the software and accelerates analysis by users beyond the FORCE primary developers. Reliability of the FORCE ecosystem is established according to the American Nuclear Society?s Nuclear Quality Assurance (NQA-1) program [American Society of Mechanical Engineers 1982], with specific levels of software quality assurance (SQA) within NQA-1 applied to each software tool in FORCE. As the tools within FORCE have matured, some integration algorithms to accurately connect the software tools for holistic analysis have been developed and deployed within the FORCE software repository. In accordance with NQA-1 standards, regression tests are required to guarantee the software performs consistently even when new capabilities are added to the software. In this report, we document the deployment of both unit tests, which test the consistent behavior of small pieces of the FORCE code base, as well as integration tests, which test the consistent performance of full use cases for the FORCE integration algorithms. We further document the encapsulation of these tests within a test harness, which collectively checks for each successful test completion on demand. Finally, we document the automation of the test harness using GitHub Actions [GitHub 2024], which require all tests succeed before any new capability or other changes can be added to the FORCE integration software

97 MATHEMATICS AND COMPUTING↗

nmrrr : A Reproducible Workflow for Binning and Visualizing NMR Spectra From Environmental Samples

Nuclear magnetic resonance (NMR) spectroscopy is a useful tool for detection and identification of molecular structural information, with increasing applications in environmental sciences. NMR instrument outputs are however heterogeneous and require extensive post-processing, creating barriers to their use and application by non-specialists. Here, we report on a new open-source R package, nmrrr, that processes and visualizes spectral data obtained from one-dimensional solution-state and solid-state NMR experiments; the package also performs relevant calculations commonly applied in natural organic matter communities, such as computing the relative abundance of various functional groups. We document the package's installation, dependencies, and functions; and provide a standard workflow for processing NMR data. This package is currently available on CRAN and GitHub, and community contributions are welcome.

54 ENVIRONMENTAL SCIENCES↗

Pilgrim Hot Springs: GEOPHIRES Inputs and Outputs for Direct-Use Geothermal District Heating and Cooling

This dataset includes files for a techno-economic analysis conducted using the GEOPHIRES simulator to examine the feasibility of expanding a larger district heating site in a remote location: Pilgrim Hot Springs, Alaska. Files included here are GEOPHIRES inputs and outputs for five different scenarios with varying demand, cycle, and system design characteristics to analyze. Also included is the link to the GEOPHIRES GitHub, as well as a link to the dataset that contains the energy modelling used to determine the heating demand for the district. For a list of the differences between scenarios, see the included "Input Overview.txt" file. Fields included in the input files are: subsurface technical parameters, surface technical parameters, financial parameters, capital and O&M parameters, as well as simulation parameters. The output files are case reports that summarize all equipment, reservoir characteristics, costs, and heating profiles.

15 GEOTHERMAL ENERGY↗

Renewable Energy Potential Model: Priority Geothermal Leasing Areas ReEDs Results

This dataset contains the results of a study conducted by the National Renewable Energy Laboratory (NREL) to identify potential future priority geothermal leasing areas on Bureau of Land Management (BLM) and United States Forest Service (USFS) lands. The analysis uses the Regional Energy Deployment System (ReEDS) model to evaluate geothermal resource potential under different scenarios of resource depth and technology combinations through the year 2050. The study considers geothermal resource potential, natural resource conflicts, and transmission access to categorize areas into near, mid, and far deployment priorities. The dataset includes outputs from the ReEDS model, such as geothermal capacity, generation, system costs, and emissions under various economic and technical scenarios. Favorability site data with geographic coordinates and site-specific attributes (e.g., resource favorability, land type) are also provided. Supporting resources include a technical report detailing methodologies and assumptions, along with a link to the ReEDS model GitHub repository, which requires GAMS and Python software for execution.

15 GEOTHERMAL ENERGY↗

TEAMER: Triton Systems Oscillating Water Column Modeling Data and Report

This dataset provides the output of six Wave Energy Converter Simulator (WEC-Sim) simulations and accompanying documentation for the modeling of Triton Systems' oscillating water column (OWC) system at tank scale (validated using available data for tuning the model, Tests 1-2) and deployment scale (for which no validation data is available, Tests 4-6). Included are the output data in a MATLAB file structure, a comprehensive report on the modeling and design of the Triton OWC system, and a link to the WEC-Sim GitHub page. This work was supported by funding from TEAMER RFTS 5 (Request for Technical Support).

16 TIDAL AND WAVE POWER↗

UNIFI's Grid-Forming (GFM) Inverter Reference Design: A Tutorial on Modeling, Control, and Experimental Implementation of GFM Inverters

The UNIFI Consortium's tutorial on grid-forming (GFM) inverters provides a comprehensive guide to the modeling, control, and experimental implementation of GFM inverters. As the integration of renewable energy accelerates, the transition from traditional grid following (GFL) to GFM inverters is crucial to ensure stable and sustainable power systems. This document outlines a reference design for three-phase and single-phase GFM inverters developed at the University of Texas at Austin. The tutorial also provides step-by-step guidance for accessing and using UNIFI’s GitHub repository, enabling users to design, build, and test GFM inverters efficiently. By fostering collaboration and equipping users with accessible resources, this initiative aims to drive widespread adoption of GFM technology across academia, utilities, and industries.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Catalyst Design in Nitrate Removal

Based on the volcano plot developed by Dr. Goldsmith group (Report linked in submission), we utilized DFT (density functional theory) calculations to search for bimetallic materials in the application of catalysts in aqueous nitrate removal. The calculations are conducted via the high-throughput automated workflow package developed by our group (Github linked in submission) using VASP commercial first-principles calculation software.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Large language model evaluation for high–performance computing software development

We apply AI-assisted large language model (LLM) capabilities of GPT-3 targeting high-performance computing (HPC) kernels for (i) code generation, and (ii) auto-parallelization of serial code in C ++, Fortran, Python and Julia. Our scope includes the following fundamental numerical kernels: AXPY, GEMV, GEMM, SpMV, Jacobi Stencil, and CG, and language/programming models: (1) C++ (e.g., OpenMP [including offload], OpenACC, Kokkos, SyCL, CUDA, and HIP), (2) Fortran (e.g., OpenMP [including offload] and OpenACC), (3) Python (e.g., numpy, Numba, cuPy, and pyCUDA), and (4) Julia (e.g., Threads, CUDA.jl, AMDGPU.jl, and KernelAbstractions.jl). Kernel implementations are generated using GitHub Copilot capabilities powered by the GPT-based OpenAI Codex available in Visual Studio Code given simple + + prompt variants. To quantify and compare the generated results, we propose a proficiency metric around the initial 10 suggestions given for each prompt. For auto-parallelization, we use ChatGPT interactively giving simple prompts as in a dialogue with another human including simple “prompt engineering” follow ups. Results suggest that correct outputs for C++ correlate with the adoption and maturity of programming models. For example, OpenMP and CUDA score really high, whereas HIP is still lacking. We found that prompts from either a targeted language such as Fortran or the more general-purpose Python can benefit from adding language keywords, while Julia prompts perform acceptably well for its Threads and CUDA.jl programming models. Finally, we expect to provide an initial quantifiable point of reference for code generation in each programming model using a state-of-the-art LLM. Overall, understanding the convergence of LLMs, AI, and HPC is crucial due to its rapidly evolving nature and how it is redefining human-computer interactions.

97 MATHEMATICS AND COMPUTING↗