Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “pre-processing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 145 records · Page 8

Quantitatively Monitoring Bubble-Flow at a Seep Site Offshore Oregon: Field Trials and Methodological Advances for Parallel Optical and Hydroacoustical Measurements

Two lander-based devices, the Bubble-Box and GasQuant-II, were used to investigate the spatial and temporal variability and total gas flow rates of a seep area offshore Oregon, United States. The Bubble-Box is a stereo camera–equipped lander that records bubbles inside a rising corridor with 80 Hz, allowing for automated image analyses of bubble size distributions and rising speeds. GasQuant is a hydroacoustic lander using a horizontally oriented multibeam swath to record the backscatter intensity of bubble streams passing the swath plain. The experimental set up at the Astoria Canyon site at a water depth of about 500 m aimed at calibrating the hydroacoustic GasQuant data with the visual Bubble-Box data for a spatial and temporal flow rate quantification of the site. For about 90 h in total, both systems were deployed simultaneously and pressure and temperature data were recorded using a CTD as well. Detailed image analyses show a Gaussian-like bubble size distribution of bubbles with a radius of 0.6–6 mm (mean 2.5 mm, std. dev. 0.25 mm); this is very similar to other measurements reported in the literature. Rising speeds ranged from 15 to 37 cm/s between 1- and 5-mm bubble sizes and are thus, in parts, slightly faster than reported elsewhere. Bubble sizes and calculated flow rates are rather constant over time at the two monitored bubble streams. Flow rates of these individual bubble streams are in the range of 544–1,278 mm 3 /s. One Bubble-Box data set was used to calibrate the acoustic backscatter response of the GasQuant data, enabling us to calculate a flow rate of the ensonified seep area (~1,700 m 2 ) that ranged from 4.98 to 8.33 L/min (5.38 × 10 6 to 9.01 × 10 6 CH 4 mol/year). Such flow rates are common for seep areas of similar size, and as such, this location is classified as a normally active seep area. For deriving these acoustically based flow rates, the detailed data pre-processing considered echogram gridding methods of the swath data and bubble responses at the respective water depth. The described method uses the inverse gas flow quantification approach and gives an in-depth example of the benefits of using acoustic and optical methods in tandem.

54 ENVIRONMENTAL SCIENCES↗

Open Data and Deep Semantic Segmentation for Automated Extraction of Building Footprints

Advances in machine learning and computer vision, combined with increased access to unstructured data (e.g., images and text), have created an opportunity for automated extraction of building characteristics, cost-effectively, and at scale. These characteristics are relevant to a variety of urban and energy applications, yet are time consuming and costly to acquire with today’s manual methods. Several recent research studies have shown that in comparison to more traditional methods that are based on features engineering approach, an end-to-end learning approach based on deep learning algorithms significantly improved the accuracy of automatic building footprint extraction from remote sensing images. However, these studies used limited benchmark datasets that have been carefully curated and labeled. How the accuracy of these deep learning-based approach holds when using less curated training data has not received enough attention. The aim of this work is to leverage the openly available data to automatically generate a larger training dataset with more variability in term of regions and type of cities, which can be used to build more accurate deep learning models. In contrast to most benchmark datasets, the gathered data have not been manually curated. Thus, the training dataset is not perfectly clean in terms of remote sensing images exactly matching the ground truth building’s foot-print. A workflow that includes data pre-processing, deep learning semantic segmentation modeling, and results post-processing is introduced and applied to a dataset that include remote sensing images from 15 cities and five counties from various region of the USA, which include 8,607,677 buildings. The accuracy of the proposed approach was measured on an out of sample testing dataset corresponding to 364,000 buildings from three USA cities. The results favorably compared to those obtained from Microsoft’s recently released US building footprint dataset.

97 MATHEMATICS AND COMPUTING↗

Tensor Extraction of Latent Features (TELF)

Tensor ELF is a user-friendly parallel tensor decomposition Python toolbox that includes a suite of machine learning algorithms for CPU and GPU architectures for the analysis of sparse and dense data including utility tools for pre-processing and post-processing.

Eren, Maksim↗

pyvisco [SWR-22-30]

pyvisco is a Python library that supports the identification of Prony series parameters for linear viscoelastic materials described by a Generalized Maxwell model. The necessary material model parameters are identified by fitting a Prony series to the experimental measurement data. pyvisco allows for the identification of Prony series parameters from experimental data measured in either the frequency-domain (via Dynamic Mechanical Thermal Analysis) or time-domain (via relaxation measurements). The experimental data can be provided as raw measurement sets at different temperatures or as pre-processed master curves. An optional minimization routine is included to reduce the number of Prony elements. This routine is helpful in Finite Element simulations where reducing the computational complexity of the linear viscoelastic material models can shorten the simulation time. See also, https://pypi.org/project/pyvisco/

Springer, Martin↗

Maps of ice wedge thermokarst pool expansion from twenty-seven circumpolar survey areas

This repository includes data and code to accompany the manuscript 'Topography controls variability in circumpolar permafrost thaw pond expansion' by Abolt et al. The data include satellite imagery and derived maps of thermokarst pools from twenty-seven survey areas in North America and Siberia. The code, written in MATLAB (R2021a), contains demonstrations of the workflow for generating the maps. The demonstrations include training a generalized UNet for mapping thermokarst pools using data from three survey areas, 'fine tuning' the UNet for use at a specific survey area using transfer learning, applying a trained UNet to infer thermokarst pool extent within satellite imagery, and performing histogram matching as a pre-processing step to improve satellite imagery contrast. Contains MATLAB script files and M files, TIF files, shape files, XML, Excel, TXT, and CSV files.The Next-Generation Ecosystem Experiments: Arctic (NGEE Arctic) was a research effort to reduce uncertainty in Earth System Models by developing a predictive understanding of carbon-rich Arctic ecosystems and feedbacks to climate. NGEE Arctic was supported by the Department of Energy's Office of Biological and Environmental Research. The NGEE Arctic project had two field research sites: 1) located within the Arctic polygonal tundra coastal region on the Barrow Environmental Observatory (BEO) and the North Slope near Utqiagvik (Barrow), Alaska and 2) multiple areas on the discontinuous permafrost region of the Seward Peninsula north of Nome, Alaska. Through observations, experiments, and synthesis with existing datasets, NGEE Arctic provided an enhanced knowledge base for multi-scale modeling and contributed to improved process representation at global pan-Arctic scales within the Department of Energy's Earth system Model (the Energy Exascale Earth System Model, or E3SM), and specifically within the E3SM Land Model component (ELM).

54 ENVIRONMENTAL SCIENCES↗

DEM, DSM, and Cleaned LiDAR Point Cloud Data from the NGEE Arctic UAS Campaigns at the Teller 27 Field Site from 2017 and 2018, Seward Peninsula, Alaska

A Digital Elevation Model (DEM) and Digital Surface Model (DSM) were derived from airborne Light Detection and Ranging (LiDAR) data collected from Los Alamos National Laboratory's (LANL) heavy-lift unoccupied aerial system (UAS) quadcopter and hexacopter platforms operated by Next-Generation Ecosystem Experiments: Arctic (NGEE Arctic) scientists from the EES-14 group at LANL. These data were collected in August 2017 and July 2018 at the NGEE Arctic field site near mile marker 27 of the Bob Blodgett Nome-Teller Memorial Highway between Nome, Alaska and Teller, Alaska. A Vulcan Raven X8 Airframe (Mitcheldean, Gloucestershire, UK), DJI Matrice 600 Pro Airframe (Shenzhen, China), and Routescene UAV LiDARSystem (Edinburgh, Scotland, UK) were used to collect LiDAR data. Following pre-processing in Routescene LidarViewer Pro software, the LiDAR point clouds were cleaned and processed using CloudCompare software to separate ground and off-ground points. A high resolution DEM and DSM were then created using ArcGIS Pro software. This data package contains fully cleaned point clouds of ground and off-ground points (.las), a 25 cm DEM (.tif), and a 25 cm DSM (.tif) for the Teller 27 field site. Ancillary aircraft data, flight mission parameters, weather conditions, and raw lidar data and imagery can be found in the L0 datasets for these campaigns: NGA299 (2017) and NGA297 (2018). Minimally processed point clouds and auxiliary files can be found in the L1 dataset: NGA304 (2017 and 2018).The Next-Generation Ecosystem Experiments: Arctic (NGEE Arctic), was a 15-year research effort (2012-2027) to reduce uncertainty in Earth System Models by developing a predictive understanding of carbon-rich Arctic ecosystems and feedbacks to climate. NGEE Arctic was supported by the Department of Energy's Office of Biological and Environmental Research.The NGEE Arctic project had two field research sites: 1) located within the Arctic polygonal tundra coastal region on the Barrow Environmental Observatory (BEO) and the North Slope near Utqiagvik (Barrow), Alaska and 2) multiple areas on the discontinuous permafrost region of the Seward Peninsula north of Nome, Alaska.Through observations, experiments, and synthesis with existing datasets, NGEE Arctic provided an enhanced knowledge base for multi-scale modeling and contributed to improved process representation at global pan-Arctic scales within the Department of Energy's Earth system Model (the Energy Exascale Earth System Model, or E3SM), and specifically within the E3SM Land Model component (ELM).

54 ENVIRONMENTAL SCIENCES↗

Expanding standards in viromics: in silico evaluation of dsDNA viral genome identification, classification, and auxiliary metabolic gene curation

Viruses influence global patterns of microbial diversity and nutrient cycles. Though viral metagenomics (viromics), specifically targeting dsDNA viruses, has been critical for revealing viral roles across diverse ecosystems, its analyses differ in many ways from those used for microbes. To date, viromics benchmarking has covered read pre-processing, assembly, relative abundance, read mapping thresholds and diversity estimation, but other steps would benefit from benchmarking and standardization. Here we use in silico-generated datasets and an extensive literature survey to evaluate and highlight how dataset composition (i.e., viromes vs bulk metagenomes) and assembly fragmentation impact (i) viral contig identification tool, (ii) virus taxonomic classification, and (iii) identification and curation of auxiliary metabolic genes (AMGs). The in silico benchmarking of five commonly used virus identification tools show that gene-content-based tools consistently performed well for long (≥3 kbp) contigs, while k -mer- and blast-based tools were uniquely able to detect viruses from short (≤3 kbp) contigs. Notably, however, the performance increase of k -mer- and blast-based tools for short contigs was obtained at the cost of increased false positives (sometimes up to ~5% for virome and ~75% bulk samples), particularly when eukaryotic or mobile genetic element sequences were included in the test datasets. Furthermore, for viral classification, variously sized genome fragments were assessed using gene-sharing network analytics to quantify drop-offs in taxonomic assignments, which revealed correct assignations ranging from ~95% (whole genomes) down to ~80% (3 kbp sized genome fragments). A similar trend was also observed for other viral classification tools such as VPF-class, ViPTree and VIRIDIC, suggesting that caution is warranted when classifying short genome fragments and not full genomes. Finally, we highlight how fragmented assemblies can lead to erroneous identification of AMGs and outline a best-practices workflow to curate candidate AMGs in viral genomes assembled from metagenomes. Together, these benchmarking experiments and annotation guidelines should aid researchers seeking to best detect, classify, and characterize the myriad viruses ‘hidden’ in diverse sequence datasets.

59 BASIC BIOLOGICAL SCIENCES↗

Smart Semi-Supervised Accumulation of Large Repositories for Industrial Control Systems Device Information

Industrial Control Systems device manufacturers frequently add new features to improve their product performance. Oftentimes, these changes are mainly vendor-driven initiatives, and customers may not be aware of the full impact of these new capabilities on their cybersecurity posture. In the energy sector, this can lead to considerable dissonance between vendor-provided cybersecurity claims and a customer’s responsibility for Operation Technology cybersecurity compliance. Thus, the resulting dynamic verification burden is shifted towards the customer and may pose a significant cybersecurity risk to the energy sector landscape. We found that there is very limited research into cybersecurity auditing for Operational Technology. However, a solution is needed for vetting the vendor-supplied feature claims and their adherence to cybersecurity requirements and standards. We are presently engaged in an effort to develop such a system. This paper demonstrates one vital aspect of this effort in proposing an end-to-end framework to accumulate a large repository of ICS device information for this vetting system, curate the dataset, and conduct extensive processing. This framework is designed to use web scraping, data analytics and Natural Language Processing (NLP) techniques to identify vendor websites, automate the collection of website-accessible documents and automatically derive metadata from them for identification of product documents relevant to the repository. We have found that this automated approach to vendor identification, document extraction into a product repository, and NLP pre-processing is unique and has not been previously presented in the literature. The preliminary work shows that this is feasible and can produce reliable results with minimum supervision. Future work will be built upon this foundation in order to achieve semi-supervised vetting of device technical information – a vital capability for ensuring that vendor-claimed device cybersecurity capabilities match industry requirements.

Ameri, Kimia↗

Detecting Low Surface Brightness Galaxies with Mask R-CNN

Low surface brightness galaxies (LSBGs), galaxies that are fainter than the dark night sky, are famously difficult to detect. However, studies of these galaxies are essential to improve our understanding of the formation and evolution of low-mass galaxies. In this work, we train a deep learning model using the Mask R-CNN framework on a set of simulated LSBGs inserted into images from the Dark Energy Survey (DES) Data Release 2 (DR2). This deep learning model is combined with several conventional image pre-processing steps to develop a pipeline for the detection of LSBGs. We apply this pipeline to the full DES DR2 coadd image dataset, and preliminary results show the detection of 22 large, high-quality LSBG candidates that went undetected by conventional algorithms. Furthermore, we find that Galactic cirrus represents the largest contaminant in our resulting candidate list.

Levy, Caleb↗

TEA Modeling to Quantify Economic Implications for Biorefinery Processing of Isolated Anatomical Fractions of Corn Stover

The Feedstock-Conversion Interface Consortium (FCIC; https://www.energy.gov/sites/prod/files/2020/01/f70/beto-fcic-overview-web.pdf), a collaboration of nine national laboratory partners, seeks to understand impacts of feedstock attributes on biorefinery performance. It is hypothesized that different individual anatomical fractions of corn stover vary in composition and recalcitrance, such that processing each fraction on its own through dedicated campaigns may enable better biorefinery economics overall relative to processing the whole stover material. This presentation focuses on techno-economic analysis (TEA) modeling to quantify the yield and cost ramifications for processing isolated anatomical fractions of corn stover through a low-temperature conversion biorefinery, reflecting a biochemical processing pathway consisting of biomass deconstruction through pretreatment and enzymatic hydrolysis, sugar fermentation and upgrading to hydrocarbon fuels, and lignin upgrading to value-added coproducts. Commercial-scale process simulation and economic evaluation leveraged experimental and analytical data from FCIC researchers for conversion of whole corn stover plus three individual anatomical fractions (cobs, husks, and stalks) across key steps of the conversion process. Our assessment found encouraging potential for biorefinery economic gains that may be achieved through this approach. TEA results indicated fuel yields varying from 29-44 gallons gasoline equivalent (GGE)/dry ton for the individual anatomical fractions compared to whole stover at 34 GGE/ton, equating to minimum fuel selling prices (MFSPs) between $6.37-$10.18/GGE for the fractions versus $8.76/GGE for whole stover (when lignin is burned), or $9.15-$15.19/GGE for the fractions versus $13.11/GGE for whole stover (when lignin is upgraded to coproducts, based on current experimental performance levels). Cobs and husks demonstrated the ability to achieve the highest fuel yields and lowest MFSPs, outperforming whole stover, while stalks led to the opposite result, as a composite reflection of compositional differences and process convertibility. Notably, even when taking the weighted average of the results reflecting each anatomical fraction weighted by its corresponding makeup of corn stover, this feasibility TEA screening supports feedstock cost allowances on the order of roughly $22-$29/ton as may reflect accommodating additional biomass fractionation equipment during feedstock pre-processing upstream of the conversion biorefinery gate to separate corn stover into such constituent fractions. Or viewed differently, the weighted average MFSP for the fractions was found to be $0.31-$0.32/GGE lower than the MFSP for the whole stover basis across either lignin scenario, when maintaining a fixed biomass feedstock cost. These findings highlight favorable implications for biorefinery economics as may be achieved by moving to a staged campaign approach for processing different corn stover fractions sequentially. Further opportunities exist for future work to fill in data gaps for remaining anatomical constituents (e.g. leaves) that were not included in the initial experimental studies, though are expected to maintain similar trends.

biochemical processing pathway↗

Phonon-informed Neural Thermal Scattering (NeTS) Optimization for Crystalline Graphite and Beryllium Metal

Fast neutrons born from fission lose energy through scattering interactions in the process of slowing-down. As neutrons thermalize to the order of $k$ $b$ $T$ (where $k$ $b$ is the Boltzmann constant, and $T$ is the temperature of the medium), their de Broglie wavelength and energy approaches the order of inter-atomic spacing and quantized lattice vibrations, i.e., phonons. At thermal energies, the thermal scattering law (TSL), i.e., $S$($α, β$), captures crystal binding contributions to the total reaction rate, or cross section. This dimensionless material property describes the energy ($β$) and momentum ($α$) exchanges available in a medium. Currently, $S$($α, β$) is evaluated in the Full Law Analysis Scattering System Hub (FLASSH) code for discrete inputs and stored as ENDF/B File 7 for 0-phonon elastic (MT 2) and n-phonon inelastic (MT 4) processes. Further processing recasts $S$($α, β$) into cumulative distribution functions for sampling post-collision scattering kinematics. In practice, interpolation schemes are employed to access data between tabulated values. An improvement to this juncture of the nuclear data pipeline is supplying cross sections on-the-fly (OTF), as has been developed for the un-resolved resonance region to minimize non-physical interpolation errors. This capability may improve simulation accuracy for accident and transient analyses, where rapidly varying changes in temperature and pressure are difficult to predict beforehand. To do so, deep artificial neural networks (ANNs) can be employed which collapse non-linear, complex data into a lightweight dictionary of neural weights and biases. This has been successfully demonstrated for the hydrogen in light water $S$($α, β$) dataset in the form of a Neural Thermal Scattering (NeTS) module. In this work, the NeTS framework is extended to consider the impact of material-dependent dynamical features on optimal neural pre-processing and architecture design decisions, such as number of neurons per hidden layer, residual skip connections and neural depth. New NeTS modules for crystalline graphite and beryllium metal illuminate a novel correlation between dynamical nonlinearity and optimal neural parametrization when deploying $S$($α, β$) on-the-fly.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Method for designing a combustion system with reduced environmentally-harmful emissions

A method for designing a combustion system which emits less of at least one environmentally-harmful emission is presented. In a describing step, an injector which introduces a fuel into a combustion chamber is described via a CFD code. In a modeling step, combustion kinetics of the fuel are modeled via a pre-processing code as the fuel mixes and reacts with an oxidizer. In a first selecting step, at least one primary scalar is derived during the modeling of the combustion kinetics. In a performing step, a table look-up is performed to obtain at least one data from a look-up database based on the primary scalar. In a second selecting step, at least one secondary scalar is selected in addition to the primary scalar(s). In a specifying step, at least one chemical pathway of formation or destruction for the secondary scalar is specified via a chemistry manager wherein the secondary scalar is representative of the environmentally-harmful emission(s) of the chemical pathway(s). In a utilizing step, the data is utilized to evaluate the chemical pathway(s) to quantify the environmentally-harmful emission(s). In an identifying step, an improvement to the combustion system is identified which reduces the environmentally-harmful emission(s).

Zambon, Andrea C.↗

A Methodology for Simulating Supercritical CO2 Heat Transfer Experiments Using Machine Learning Models

To support the growth of supercritical carbon dioxide (sCO2) power cycles in the energy industry, this study seeks to train a machine learning model to mirror experimental data to predict new heat transfer data. To do this experimental data was amassed, one preliminary set comprised of 16 test results, and an expanded version comprised of 38 test results. With the goal of predicting experimental apparatus temperatures and pressures, several iterations of models were tested investigating the impact of model hyper-parameters, data inclusion, and data pre-processing on model performance. A total of 15 variations cumulatively of Gaussian Process Regressors, Gradient Boosting Regressors, and Multi-Layer Perceptrons were trained and validated on the preliminary set, and the best algorithm of each class was re-trained on the expanded set. These were compared based on test/train R^2 , test/train mean absolute error (MAE), and validation MAE, to identify the successfulness of these models. It was shown temperatures could be predicted within just a few degrees, showing the potential of this approach. Future research has been identified with approaches to improve pressure and temperature predictions going forward.

Grabowski, Owen↗

Improving and Automating Building Model Data Exchange

There are many instances throughout a project’s lifecycle where there arises a need for quick and accurate risk assessment of building designs. For example, an unexpected design change during construction may necessitate structural engineers to perform a seismic risk assessment on analytical models of the updated building design using high fidelity structural analysis software, such as ANSYS or Abaqus. However, the efficiency of such workflows often depends upon the interoperability of architectural design software and structural analysis software. When the quality of this interoperability is lacking or even non-existent, the efficiency of virtual engineering workflows is hampered, which increases project costs. A McGraw Hill industry survey of professional users of Building Information Modeling (BIM) technologies found that there is high demand for BIM interoperability for structural analysis, but that the value/difficulty ratio is currently too low for practical use. There have been efforts by the academic community to facilitate model data exchange between the architectural design and structural analysis domains, but such solutions have not been widely adopted by industry, face technical challenges, and oftentimes are limited in applicability for users of various BIM software. Therefore, INL is developing capabilities to improve, automate, and generalize model data exchange between architectural BIM software (e.g., Revit) and structural analysis software (e.g., SAP2000, ANSYS). The goal is to help expedite and automate as much of the pre-processing step for creating analytical models in finite element analysis software as reasonably as possible. Such a "BIM-to-FEA" conversion tool should provide direct benefit to end-users through accuracy, automation, quick turn-around, and wide applicability. To generalize the application of this BIM-to-FEA conversion tool and increase its useability among the many different commercial BIM software currently used by industry, the program is being developed with the concept of openBIM. OpenBIM is the application of non-proprietary, open data standards that allow for BIM model data exchange in a format that is accessible, retainable, and useable for all users. The most widely used open, non-proprietary data exchange format for BIM is the Industry Foundation Classes (IFC) schema. IFC is developed by buildingSMART international and is ISO certified (ISO 16739-1:2018). The BIM-to-FEA conversion tool is being developed for compatibility with typical commercial building designs of steel framed structures. The tool is currently capable of importing architectural BIM data of framed building structures, recognizing and extracting the aspects of the model that are required for structural analysis, adjusting the connectivity of frame members, and finally exporting to an analytical model stored in the IFC format. The exported IFC analytical model can then be imported into various openBIM compliant software, such as SAP2000. Such capabilities have already been tested on commercial software, as shown above, and continue to be improved. Work is underway to test the conversion on various commercial BIM software, develop a user-friendly interface, incorporate the program into the broader DeepLynx data warehouse project being developed by INL, and to eventually open-source the tool for the benefit of the community. Future development of the tool envisions the ability for efficient iterative risk assessment of generative building designs, all within a workflow utilizing open-source tools. One such open-source tool will be MOOSE, an advanced finite element analysis tool developed at INL. The conversion tool will also branch out from typical commercial building designs and will aim to incorporate nuclear construction. The aim will be to convert both structural and non-structural components of nuclear facilities, such as curved concrete containment structures and piping systems, respectively.

97 MATHEMATICS AND COMPUTING↗

System Engineers and Decisions: It?s All about Knowledge

In order to guarantee that a system meets adequate levels of reliability and availability, system performances are continuously monitored and analyzed thanks to the technological advancements driving the Industry 4.0 revolution. An Industry 4.0 approach is typically based on advanced statistical, big data mining, machine learning, and internet-of-things methods designed to detect anomalies in the behavior of system, detect the most likely failure modes, and provide indications to system engineers on when maintenance activities should be performed before system performance are deemed unacceptable (which can be generated by diagnostic and prognostic methods). However, these analyses, which are designed to automatize and increase the efficacy of the system maintenance program, require large amount of data which can come in various forms: numeric, textual, images, sounds etc. Such data constitutes the historic knowledge benchmark to track system performances and support system engineer decisions. Here we claim that data is not sufficient to support this kind of analyses when applied to systems characterized by complex architectures and behaviors. Robust system engineer decisions require the ability to understand the system operational context that lies behind the observed data elements. In this respect, system models are in fact necessary to “put data in context” and capture relationships between data elements. Industry 4.0 methods require in fact contextual knowledge as a basis upon which hypotheses can be generated and assumptions tested. In our view, for complex systems, model-based system engineering (MBSE) models can afford this contextual knowledge, as they are typically used to describe systems architecture and dynamic behaviors. System knowledge is here intended as the blending of collected data and system architecture which takes the form of a “knowledge graph”. A knowledge graph is a database which consists of a large set of nodes (in our case an entity can be either a data or an MBSE element) which are linked to each other. The types of nodes and links follow a pre-defined topology, sometimes also refers as an ontology, that is designed to fit the actual decisions that needs to be performed. We show here how a knowledge graph can be defined to support system engineer maintenance decisions and how the same graph can be built based on system MBSE models and pre-processed data from numeric (through anomaly detections and diagnostic methods) and textual elements (through technical language processing TLP).

97 - MATHEMATICS AND COMPUTING↗

Impact of low-chemical storage pretreatment of loblolly pine bark on biochar from microwave pyrolysis

Forest product residues such as bark represent a low-cost, abundant feedstock for bioenergy, but their high ash and alkali and alkaline earth metal (AAEM) content limit thermochemical conversion efficiency. This study evaluates the use of low-severity chemical pretreatments during anaerobic storage to improve the performance of microwave pyrolysis for loblolly pine bark. Bark was treated with dilute sulfuric acid (0.1% and 1%, w/w) or sodium hydroxide (4%, w/w) and incubated anaerobically for one or two weeks to simulate in-pile biorefinery storage. The most effective treatment—1% H2SO4 for two weeks—reduced AAEM content by 35.7% and increased bio-oil yield by 11% compared to untreated controls, while also reducing pyrolysis gas production. In contrast, alkali treatment did not reduce AAEM levels and led to decreased bio-oil yields with increased gas formation. Although biochar yields were relatively stable across treatments, their physicochemical characteristics varied significantly. Acid-treated bark yielded biochars with higher carbon content, lower O/C and H/C ratios, greater surface area, and enhanced heating values. These improvements suggest that chemical pretreatment during storage can tailor biochar quality for specific end uses. Biochars produced under optimized conditions exhibited properties suitable for soil amendment, carbon sequestration, and solid fuel applications. This integrated approach—combining storage, mild chemical conditioning, and microwave pyrolysis—provides a viable pathway to enhance the value and sustainability of bark-derived bioenergy products.

09 - BIOMASS FUELS↗

Analytical Modeling of Biomass Transport and Feeding Systems

The processing of biomass solids in a biorefinery consists of pretreatment, enzyme hydrolysis / concurrent fermentation of sugars to ethanol, product recovery, and drying. Sustainable operation requires a front end that transforms wet solids into a pumpable slurry. Otherwise the biorefinery will suffer unscheduled shut-downs and inefficient operation due to solids that obstruct pumps and other equipment and resist mixing in a bioreactor. Downtime in pioneer biorefineries due to interruptions from materials handling problems has been 50% or more, leading to unsustainable manufacturing processes. This work addresses new technology, predictive computational models, and definition of operational conditions that result in formation of slurries of corn stover at up to 300 g/L using low enzyme loadings (1 to 3 FPU cellulase/g) before the biomass (corn stover) enters the pretreatment step. A team of researchers from Purdue University, Idaho National Laboratory (INL), Forest Concepts, AdvanceBio, Argonne National Laboratory, and DOE BETO have combined their knowledge in agricultural and biological engineering, bioprocess engineering, mechanical engineering, chemical engineering, agricultural economics, materials engineering and enzyme and microbial technology to address the challenge of making lignocellulose flow. This team effort has resulted in the development and validation of conditions that employ low levels of commercial enzyme in an agitated bioreactor to which corn stover pellets are added resulting in formation of slurries at high solids loadings, before pretreatment. This approach overcomes challenges caused by handling of dry, particulate biomass materials at the front end of the biorefinery. The subsequent materials handling issues cause obstruction at pumps, pipes and valves. Formation of high loadings slurries with low yield stress, as reported here, significantly decreases the potential for process interruption and enhances plant operability. Key advances in the knowledge of how slurry formation occurs is reported here and in recently published journal papers. We found that pellets are needed to achieve high solids loading, and that commercial enzymes are effective in forming slurries of corn stover particles from pellets that have not been pretreated. Our work has resulted in models that predict solids behavior for formation of compressed solids and pellets that in turn facilitate slurries made of high concentrations of corn stover particles. A computational model was developed that gives mechanistic insights into properties of particles and mixing process that gives the slurry rheology needed to facilitate pumping. Hence, the corn stover may be pumped into a pretreatment reactor in place of auguring in solids against high pressure which is a root cause of interruptions at the front end of a biorefinery. Subsequent mixing in enzyme and microbial bioreactors results in conversion of lignocellulose to sugars in a biorefinery in agitated bioreactors, with flows in and out of the vessels being less likely to be interrupted due to plugging or materials handling problems. The obtained data coupled to process models, techno-economic assessment (TEA) and Life Cycle Analysis (LCA) were used to assess whether this approach is practical. These results are based on a foundation of laboratory characterization and pilot runs. The NREL biochemical sugar model was utilized to carry out techno-economic analysis of enzyme catalyzed liquefaction followed by enzyme hydrolysis. The minimum sugar selling price was between 17.5 and 18.3 ¢/pound or about the same as calculated by the NREL model for dilute acid pretreatment followed by enzyme hydrolysis. Life cycle analysis (LCA) based on Argonne’s Greet Model showed the enzyme catalyzed route had the lowest greenhouse gas emissions of the three combinations studied (i.e., enzyme, enzyme mimetic, and enzyme + mimetic combined). GHG emissions for enzyme-based corn stover liquefaction step, alone, were about 21 g CO 2 -equivalent/kg of liquefied slurry. We believe this approach will further enhance operability of a pioneer biorefinery, and bring large-scale conversion of lignocellulosic biomass to low carbon footprint biofuels closer to implementation.

09 BIOMASS FUELS↗

Near-Infrared Spectroscopy can Predict Anatomical Abundance in Corn Stover

Feedstock heterogeneity is a key challenge impacting the deconstruction and conversion of herbaceous lignocellulosic biomass to biobased fuels, chemicals, and materials. Upstream processing to homogenize biomass feedstock streams into their anatomical components via air classification allows for a more tailored approach to subsequent mechanical and chemical processing. Here, we show that differing corn stover anatomical tissues respond differently to pretreatment and enzymatic hydrolysis and therefore, a one-size-fits-all approach to chemical processing biomass is inappropriate. To inform on-line downstream processing, a robust and high-throughput analytical technique is needed to quantitatively characterize the separated biomass. Predictive correlation of near-infrared spectra to biomass chemical composition is such a technique. Here, we demonstrate the capability of models developed using an “off-the-shelf,” industrially relevant spectrometer with limited spectral range to make strong predictions of both cell wall chemical composition and the relative abundance of anatomical components of the corn stover, the latter for the first time ever. Gaussian process regression (GPR) yields stronger correlations (average R 2 v = 88% for chemical composition and 95% for anatomical relative abundance) than the more commonly used partial least squares (PLS) regression (average R 2 v = 84% for chemical composition and 92% for anatomical relative abundance). In nearly all cases, both GPR and PLS outperform models generated using neural networks. These results highlight the potential for coupling NIRS with predictive models based on GPR due to the potential to yield more robust correlations.

09 BIOMASS FUELS↗