Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Parsing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Variation in Root Exudate Composition Influences Soil Microbiome Membership and Function

Root exudation is one of the primary processes that mediate interactions between plant roots, microorganisms, and the soil matrix, yet the mechanisms by which exudation alters microbial metabolism in soils have been challenging to unravel. Here, utilizing distinct sorghum genotypes, we characterized the chemical heterogeneity between root exudates and the effects of that variability on soil microbial membership and metabolism. Distinct exudate chemical profiles were quantified and used to formulate synthetic root exudate treatments: a high-organic-acid treatment (HOT) and a high-sugar treatment (HST). To parse the response of the soil microbiome to different exudate regimens, laboratory soil reactors were amended with these root exudate treatments as well as a nonexudate control. Amplicon sequencing of the 16S rRNA gene illustrated distinct microbial diversity patterns and membership in response to HST, HOT, or control amendments. Exometabolite changes reflected these microbial community changes, and we observed enrichment of organic and amino acids, as well as possible phytohormones in the HST relative to the HOT and control. Linking the metabolic capacity of metagenome-assembled genomes in the HST to the exometabolite patterns, we identified microorganisms that could produce these phytohormones. Our findings emphasize the tractability of high-resolution multiomics tools to investigate soil microbiomes, opening the possibility of manipulating native microbial communities to improve specific soil microbial functions and enhance crop production.

59 BASIC BIOLOGICAL SCIENCES↗

New insights into the natural history of bronchopulmonary dysplasia from proteomics and multiplexed immunohistochemistry

Bronchopulmonary dysplasia (BPD) is a disease of prematurity related to the arrest of normal lung development. The objective of this study was to better understand how proteome modulation and cell-type shifts are noted in BPD pathology. Pediatric human donors aged 1–3 yr were classified based on history of prematurity and histopathology consistent with “healed” BPD (hBPD, n = 3) and “established” BPD (eBPD, n = 3) compared with respective full-term born (n = 6) age-matched term controls. Proteins were quantified by tandem mass spectroscopy with selected Western blot validations. Multiplexed immunofluorescence (MxIF) microscopy was performed on lung sections to enumerate cell types. Protein abundances and MxIF cell frequencies were compared among groups using ANOVA. Cell type and ontology enrichment were performed using an in-house tool and/or EnrichR. Proteomics detected 5,746 unique proteins, 186 upregulated and 534 downregulated, in eBPD versus control with fewer proteins differentially abundant in hBPD as compared with age-matched term controls. Cell-type enrichment suggested a loss of alveolar type I, alveolar type II, endothelial/capillary, and lymphatics, and an increase in smooth muscle and fibroblasts consistent with MxIF. Histochemistry and Western analysis also supported predictions of upregulated ferroptosis in eBPD versus control. Finally, several extracellular matrix components mapping to angiogenesis signaling pathways were altered in eBPD. Despite clear parsing by protein abundance, comparative MxIF analysis confirms phenotypic variability in BPD. This work provides the first demonstration of tandem mass spectrometry and multiplexed molecular analysis of human lung tissue for critical elucidation of BPD trajectory-defining factors into early childhood.

60 APPLIED LIFE SCIENCES↗

Archparse

Archparse is a Python package that holds the purpose and capability of converting the contents of a text file to a functioning neural network based on the Tensorflow 2.X framework. This is able to be done with minimal written Python code and next to no knowledge of how to build models with Tensorflow directly. More specifically Archparse parses a text file with extension ".arch" which contains neural network architecture information that corresponds to either the Tensorflow 2.X API or custom code written with the Tensorflow 2.X API. Included in the initial version is the capacity to easily produce sequential autoencoders and sequential neural networks.

Vander Wal, MichaelD.↗

Link Parser Library

Link Parser Library is a Python library for parsing HTTP Link header format. Link Parser extracts information about relationships to other resources from Link Header and translates it into an easy to use Json structure.

Balakireva, Lyudmila↗

Reposcanner

SAND2023-05455O Reposcanner provides a highly modular, extensible framework for defining routines for mining data from software repositories and performing analyses on that data to yield valuable insights on team behaviors. Reposcanner features seamless support for different version control platforms like GitHub, Gitlab, and Bitbucket; smart parsing of URLs; intelligent credential management capabilities; and a comprehensive test suite. Reposcanner is connected to the Exascale Computing Project and is intended for research purposes.

Mundt, Miranda↗

EMF Biomass Data Interpreter [SWR-23-06]

EMF Biomass Data Interpreter was developed for the transportation team of the Energy Modeling Forum (EMF) 37: Deep Decarbonization and High Electrification Scenarios for North America project. It parses results from different modeling groups to show how biomass might be employed under electrification.

Wachs, Elizabeth↗

Parsnip Parser Creation Application

Parsnip has three parts. The first part is the front-end user experience. The front end will be a graphical representation of the intermediate language. Once a user is satisfied with the information on the front end, Parsnip translates the data from the visual application into the second part of Parsnip - the intermediate language. More advanced users may skip the front end and generate their own intermediate language files. The final part is the backend which takes the intermediate language files and generates Zeek and Spicy code. Parsnip will not completely replace parser developers. Many protocols have unique challenges requiring manual effort; however, the goal of Parsnip is to automate at least 90% of the development that largely consists of repetitive tasks. Parsnip output will compile a functioning parser but may not include all PDU types or parse all data.

Huddleston, TimothyA.↗

CIMantic Graphs

CIMantic Graphs (aka CIM-Graph) is a new python library developed by PNNL to reduce the burden of working with the Common Information Model. CIMantic Graphs takes a novel approach of building in-memory labeled property graphs for creating, parsing, and editing CIM power system models.

Anderson, Alexander↗

RadSim: ENSDF, Xray and N42

This package includes three parts, (1) gov.bnl.nndc.ensdf, (2) gov.nist.xray and (3) gov.nist.physics.n42, which are utilized in the development of Radiation Detector Simulator (RadSim) project. RadSim is being developed to provide the capability to: (1) simulate radiation source emissions, (2) interpolate results from radiation transport tools into a common format to prepare incident flux, and (3) model radiation detector response to rapidly produce synthetic radiation measurement templates. RadSim is targeted for open-source release, which will enable researchers and industry partners to model gamma-ray detectors response to simulated flux from the transport tool of their choice. The first tool of the package, gov.bnl.nndc.ensdf, includes functionality to parse and split publicly available decay records in ENSDF format, and pull the relevant information from ENSDF libraries. The second tool, gov.nist.xray, provides Xray information using the public NIST database. Lastly, gov.nist.physics.n42, is a tool used to convert an N42 xml file into a Java object that can be interfaced within a Java program.

Hangal, DnaushA↗

FlowBench Raw Data Archive

The repo that provides data archive for DOE PoSeiDon project. It also contains scripts and instructions to parse the data.

George, Papadimitriou↗

Wind Energy Intrusion Detection System

SAND2022-14659 O The Wind Energy Intrusion Detection System uses a human-machine interface (HMI) to parse results and analyze for anomalies. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Johnson, Jay↗

Replete v. 2.0

SAND2021-15136 O Replete is a Java library that contains several common and useful utilities in a variety of categories including expression parsing, multi-threading, data pipelines, user interface, database connectivity, Unix shell emulation, information generalization, and Maven standardization. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

McClain, Jonathan↗

A Data Processing Pipeline To Extract A Knowledge Graph From Sec Documents For Socio-technical Analysis Of Critical Infrastructure Influence

The code is written in Python and consists of the following pipeline that is implemented in Apache Airflow. This pipeline intends to understand the companies that are directly or indirectly involved with a type of critical infrastructure system at some point in that system's lifecycle. The pipeline takes a configuration file that specifies a list of initial companies to consider, a geographic region of interest (disk) expressed as a latitude/longitude point and distance, and a set of SEC form types from which to extract entities and relations. There are three main components to this pipeline as currently implemented: Social Network Extraction, Critical Infrastructure Network Extraction, and Inference and Fusion. First, Social Network Extraction, implemented as the `organizations_sec` component of the workflow graph queries the SEC EDGAR webservice using the list of initial companies from the configuration file. Given this, it extracts metadata that documents the number of each type of form for the given set of companies and their location. This forms metadata represents a catalog of data sources for the extracted social network knowledge graph. The pipeline then downloads these forms from the website and saves them in a build directory for further processing. These documents are then parsed for entities and relations. Second, the Critical Network Extraction component extracts entities and relations for a critical infrastructure sector. Currently, we focus on Electric Vehicle charging stations and this information is available via the Department of Energy (DOE) database on fueling stations maintained by NREL. Third, the Inference and Fusion component relates the social network graph to the critical infrastructure graph in order to understand the impact of a company within a geographic region. Relations include ownership of the EV Charging Station asset as well as maintenance/ownership of the EV payment networks. The fused network can be represented in many ways and currently we emit a knowledge graph.

Weaver, GabrielA.↗

lanl-ansi/MG-RAVENS

The MG-RAVENS project with the DOE Office of Electricity Microgrid R&D Program is a project to develop a completely free, open-source data exchange standard (API) for the Department of Energy, targeted at software tools related to infrastructure modeling, particularly the modeling of microgrids and electric power distribution systems that are created with funding from the Microgrid R&D Program. This software produces formal definitions of an API, documentation, contains supporting functions for parsing, validating, etc., and will contain examples of workflows enabled by the developed API.

Fobes, David M↗

Library-AI-Toolset

Collection of tools designed to parse documents, such as PDFs, and extract structured elements including URLs, citation contexts, tables, formulas, and figures. This toolset leverages AI-based text extraction and classification methods, providing robust solutions for various scholarly resources processing needs.

Balakireva, Lyudmila↗

EyeON

EyeON: Eye on Operational technology Software Supply Chain attacks have risen drastically over the past few years, none more well-known and impactful than the SolarWinds compromise. Criminal organizations inserted an attack vector into a specific version of the source code, giving themselves an air of credibility. Once news broke on SolarWinds, identifying compromised sites was very difficult, even knowing the culprit update. Software Bills of Materials (SBOM) have been touted as the solution to reclaiming control of your software supply chain. Deployment of SBOMs has been slow, however, due to conflicting standards, opaque storage requirements, and vendor adoption. Additionally, the path from obtaining an SBOM and securing your supply chain is unclear; how can an SBOM library be leveraged to provide insight to your attack surface? The EyeON tool, sponsored by Department of Energy Cybersecurity, Energy Security, and Emergency Response (DoE CESER), aims to address these gaps by providing an encapsulated solution to tracking which updates have been installed in an enterprise, and alerting system administrators to vulnerabilities as they become known. Similar to a virus scanner, EyeON is a command line tool to parse either a single file, nested directory structure, or filesystem. It collects data such as signature (hashes), version information, VirusTotal tags, compiler, compilation date, and code signing information. Users will anonymously submit scan data periodically to DoE CESER, who will then compile a database of known software products employed by Critical Infrastructure and broadcast alerts based on discovered flaws as they arise.

Tenzing, Wangmo↗

test-grid-buildouts [SWR-23-112]

Test-grid-buildouts is a set of notebooks for creating buildouts of synthetic transmission grids with various levels of renewable energy. Workflows can be used to augment various test grid with varying levels of wind/solar generation. Aggregated renewable energy time series (actuals and scenarios) at grid buses can then be used for various transmission grid operation problems. Highly customizable workflow consists of parsing of synthetic grid files, setting up SAM & reV simulations, application of exclusion masks, creating of custom renewable generators on the grid, aggregating of power time series, and writing of modified grid description files.

Satkauskas, Ignas↗

pyscan-tlk

SAND2024-13867O pyscan-tlk software provides straight-forward access to control Thorlabs brand instruments with python. C bindings and python wrappers are generated and automatically based on text parsing the C documentation. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Mounce, Andrew↗