Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “multivariate analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 181 records · Page 10

Interdependence in active mobility adoption: Joint modeling and motivational spillover in walking, cycling and bike-sharing

Active mobility offers an array of physical, emotional, and social well-being benefits. However, with the proliferation of the sharing economy, new nonmotorized means of transport are entering the fold, complementing some existing mobility options while competing with others. The purpose of this research study is to investigate the adoption of three active travel modes—namely walking, cycling, and bikesharing—in a joint modeling framework. Here, the analysis is based on an adaptation of the stages of change framework, which originates from the health behavior sciences. Multivariate ordered probit modeling drawing on U.S. survey data provides well-needed insights into individuals’ preparedness to adopt multiple active modes as a function of personal, neighborhood, and psychosocial factors. The research suggests three important findings. (1) The joint model structure confirms interdependence among different active mobility choices. The strongest complementarity is found for walking and cycling adoption. (2) Each mode has a distinctive adoption path with either three or four separate stages. We discuss the implications of derived stage-thresholds and plot adoption contours for selected scenarios. (3) Psychological and neighborhood variables generate more coupling among active modes than individual and household factors. Specifically, identifying strongly with active mobility aspirations, experiences with multimodal travel, possessing better navigational skills, along with supportive local community norms are the factors that appear to drive the joint adoption decisions. This study contributes to the understanding of how decisions within the same functional domain are related and help to design policies that promote active mobility by identifying positive spillovers and joint determinants.

42 ENGINEERING↗

Large-Scale Inference of Multivariate Regression for Heavy-Tailed and Asymmetric Data

Large-scale multivariate regression is a fundamental statistical tool with a wide range of applications. Here, this study considers the problem of simultaneously testing a large number of general linear hypotheses, encompassing covariate-effect analysis, analysis of variance, and model comparisons. The challenge that accompanies a large number of tests is the ubiquitous presence of heavy-tailed and/or highly skewed measurement noise, which is the main reason for the failure of conventional least squares-based methods. For large-scale multivariate regression, we develop a set of robust inference methods to explore data features such as heavy tailedness and skewness, which are not visible to least squares methods. The new testing procedure is based on the data-adaptive Huber regression and a new covariance estimator of regression estimates. Under mild conditions, we show that our methods produce consistent estimates of the false discovery proportion. Extensive numerical experiments and an empirical study on quantitative linguistics demonstrate the advantage of the proposed method over many state-of-the-art methods when the data are generated from heavy-tailed and/or skewed distributions.

97 MATHEMATICS AND COMPUTING↗

Analysis of Inlier and Outlier Compounds with Respect to Artificial Neural Network Cetane Number Prediction Accuracy

Artificial neural networks (ANNs) are exceptional at forming non-linear correlations between multivariate input and target variables; however, they are often seen as a “black box” approach, since how ANNs form these correlations is somewhat ambiguous. Furthermore, the process underlying how ANNs learn from inlier and outlier samples within the input dataset is not fully understood. Intuitively, it is expected that training ANNs with inlier samples will increase prediction accuracy and training with outlier samples will reduce prediction accuracy; though, in practice, this is not always true. The present work identifies and analyzes inliers and outliers of existing experimental cetane number (CN) data encompassing a variety of compounds and compound groups. It also investigates how ANNs trained to predict CN perform with and without outliers included in the training data, and whether a relationship exists between inliers/outliers and ANN prediction accuracy across the whole dataset and for individual samples. Additionally, individual outlier compounds are analyzed, highlighting how they structurally differ from inlier compounds.

09 BIOMASS FUELS↗

Emissions mitigation technology for advanced water-lean solvent-based CO 2 capture processes

This technical final report submitted to DOE/NETL presents all the research activities performed during the entirety of DE-FE0031660 project-Emissions Mitigation Technology for Advanced Water-Lean Solvent-Based CO 2 Capture Processes which spans from October 2018 through March 2022. RTI International has been conducting studies from fundamental and operational aspects to reduce the overall amine emissions from the advanced Water-Lean Solvent (WLS) systems, specifically RTI’s Non-Aqueous Solvent (NAS). This technical final report will highlight the key findings from project which align closely to the project objectives which are: Identify the contribution of vapor loss, entrainment, and aerosols to the overall emissions of water-lean systems; Determine the significance of CO 2 capture system operating parameters to the amine emissions; Develop an emissions model based on critical operating parameters; Evaluate the effectiveness of emissions mitigation devices to reduce the amine emissions to <1 ppm under flue coal-fired flue gas; and, Determine the contribution of the ECTs to the overall CO 2 capture cost. The following are the key findings based on numerous tests using both lab-scale setups and parametric testing performed at RTI’s Bench-scale Gas Absorption System (BsGAS). During the BP1, the aerosol generation system and monitoring equipment were installed at BsGAS to produce and determine the aerosol characteristics during the NAS CO 2 capture process. The aerosol produced by this setup produced aerosols with the peak diameter and concentration of 50 micron and 1.2E10 7 cm -3 , respectively. These particle sizes and concentrations are matched to those observed in the actual coal-fired power plant flue gases and expected to be found at the absorber inlet of the CO 2 capture system. Over 1,300 hours of parametric testing have been conducted to evaluate the impact of the aerosols and operating conditions during the CO 2 capture with NAS on the overall amine emissions in the treated flue gas. At the worse condition tested, the presence of the aerosols in the flue gas could increase the overall emissions by 10X compared to the baseline emissions from NAS’s vapor pressure. CO 2 capture rate was found to be a main factor impacting the overall emissions as well as aerosol size and concentrations in the absorber off-gas. The higher CO 2 capture rate, the higher amine emissions in the treated gas. The temperature difference between the temperature bulge seen in the absorber and the water wash temperature also impacts the particle growth where the larger the temperature difference, the more amine emissions from aerosols in the treated gas. The majority of the aerosols did not grow substantially in the system, and the particle concentrations remained nearly constant between the absorber inlet and wash outlet. Only a small portion of the particles were found to grow significantly. The high efficiency demister with mesh size of 5-10 micron can be installed to remove a portion of the aerosols from the gas stream leaving the water wash. Overall, these results from parametric testing have established the emission baseline and validate our assumption on the need of emission control technologies (ECT) in order to minimize the emissions from the baseline NAS CO 2 capture process. Over 2,000 of BsGAS operating hours was used to investigate a handful of process improvements which led to a selection of the vital few changes that effectively control the amine emissions. These process improvements are lime-coated-filters for absorber gas inlet, advanced demister at the top of the absorber, a second water wash with amine recovery unit were designed, installed, and tested at BsGAS at the end of BP1. The result showed that the NAS CO 2 capture process with these additional emission control devices could lower the amine emission in the treated gas to about 1 ppm using a simulated coal-fire flue gas stream. The main contributor in lowering the amine emission came from the second water wash with amine recovery unit where the amine concentration in the scrubbing water was kept below 2 wt% through a continuous amine removal via an adsorbent bed, resulting in a low amine vapor pressure. The adsorbent bed was regenerated via a direct steam regeneration and the recover amine was returned to the absorber to minimize wastewater and makeup amine. A flue gas generation system was designed and installed during the first half of BP2 to support the emission testing using a real coal-derived flue gas. The system is capable of generating both coal- and natural gas- derived flue gases with the composition of the gaseous species highly resemble to that of the power plant flue gases. The particulates detected in the coal-derived flue gas showed the mean diameter of 1 micron. The CO 2 capture operating was then proceed using the real coal-derived flue gas where the amine emission was controlled to be about 0-3 ppm for the total run time of about 200 hours. Similar testing was conducted with natural gas-derived flue gas and the result showed a highly amine emission of 30 ppm under the total run time of 200 hours. The Principal Component Analysis (PCA) and the Partial Least Squares Projection to Latent Structures (PLS) techniques were applied to the parametric testing data to derive a multivariate statistical model. The model was validated and trained with half of the data collected, and the predictive ability of the model was evaluated using the remaining half of the data. The resulting empirical model was capable of predicting the overall emissions from the NAS process without the ECTs with ±15% accuracy (average absolute deviation, AAD) in BP1. As more emission data were obtained under the real coal-flue gas in the BP2, the model incorporated these new set of data to reflect the final process configuration, operating parameters, and amine emission. This results in the updated empirical model predicting the amine emission from the NAS CO 2 capture process with 84% goodness-of-fit (R 2 ), 85% predictability (Q 2 ), and 15% AAD. The study evaluates the use of RTI’s Non-Aqueous Solvent technology for 90% CO 2 capture from a net 650 MWe pulverized coal power plant, downstream of the flue-gas desulfurization unit. The captured CO 2 has a purity of > 95% CO 2 , and is dried, compressed to 15.3 MPa (2,215 psia), ready for sequestration. The analysis uses Case B12B from the DOE Baseline study on Bituminous Coal, Revision 4 where the Cansolv CO 2 capture plant is replaced by the RTI CO 2 Capture plant. The CO 2 capture plant has been sized to capture >90% CO 2 from flue gas derived from a net 650 MWe supercritical pulverized coal power plant. The CO 2 capture plant is equipped with emission control technologies that limits the amine emissions to < 1 ppm. Two different cases were evaluated for the technoeconomic study. The key difference between the two cases is the regenerator pressure. In Case 1, the regenerator operates at 0.195 MPa (28.3 psia), whereas in Case 2, the regenerator pressure is 0.44 MPa (64 psia) thus removing the need for the first stage of compression of the eight-stage compression train. Results from the TEA are compared against the DOE reference cases for SCPC plant with and without CO 2 Capture (Case B12A and Case B12B of the DOE Baseline study, respectively). Case 2 with CO 2 regeneration at higher pressure results in the lower cost of CO 2 capture. The total capital cost of the capture process has been estimated using 2018 dollars in Aspen Process Economic Analyzer and was estimated to be $579 MM. The capture plant operation leads to a total parasitic power loss rate of 96 MWe, resulting in a decrease in pulverized coal power plant efficiency of 7.8% points. The resulting cost of electric power increases from 64.4 mills/kWh, for no capture, to 97.5 mills/kWh, with 90% capture, an increase of 51% in the COE. The cost of capturing 90% CO 2 was estimated to be $38.2/tonne-CO 2 , and meets the DOE target of $40/t-CO 2 . Emission control technologies (ECT) investigated in this project includes a second water wash with use of activated carbon beds for removal of amine from the wash water prior to recirculation in the water wash. These ECT allow operation of the CO 2 capture plant with < 1 ppm amine emissions with the treated flue gas and contributes to $2.4/t-CO 2 captured. Amine emissions derived from thermal and oxidative degradations were investigated under this project along with the emissions derived from aerosols for the NAS system. The thermally degraded of the lean NAS showed less than 4% decreased of the original total amine content in the NAS at 150 °C while the result obtained at 120 °C showed no drop in total amine content, suggesting that thermal degradation of the NAS is minimal. These results also suggested that the thermally degraded species are not likely formed and contributed to the emissions due to the low regeneration temperature of the NAS at 90-105 °C. The oxidative degradation, on the other hand, could become problematic as some of these oxidative degraded species were observed during the NAS-5 testing at National Carbon Capture Center (NCCC) and SINTEF in our previous project. The rapid screening of selected inhibitors suggested that oxidative degradation of NAS can be suppressed using thiol containing compounds in amounts of at least 1 mol%. The detailed mechanistic degradation pathway was conceived for a specific amine used in NAS formulation during BP2. he reduction of the nitrosamines caused by the NO x present in the flue gas was also examined. The study suggested that the thermo-chemical treatment of the NAS solvent would be a more effective and economically viable compared to removing NO x at the DCC.

01 COAL, LIGNITE, AND PEAT↗

Quantitative Representativeness and Constituency of the Long-Term Agroecosystem Research Network and Analysis of Complementarity with Existing Ecological Networks

Abstract Studies conducted at sites across ecological research networks usually strive to scale their results to larger areas, trying to reach conclusions that are valid throughout larger enclosing regions. Network representativeness and constituency can show how well conditions at sampling locations represent conditions also found elsewhere and can be used to help scale-up results over larger regions. Multivariate statistical methods have been used to design networks and select sites that optimize regional representation, thereby maximizing the value of datasets and research. However, in networks created from already established sites, an immediate challenge is to understand how well existing sites represent the range of environments in the whole area of interest. We performed an analysis to show how well sites in the USDA Long-Term Agroecosystem Research (LTAR) Network represent all agricultural working lands within the conterminous United States (CONUS). Our analysis of 18 LTAR sites, based on 15 climatic and edaphic characteristics, produced maps of representativeness and constituency. Representativeness of the LTAR sites was quantified through an exhaustive pairwise Euclidean distance calculation in multivariate space, between the locations of experiments within each LTAR site and every 1 km cell across the CONUS. Network representativeness is from the perspective of all CONUS locations, but we also considered the perspective from each LTAR site. For every LTAR site, we identified the region that is best represented by that particular site—its constituency—as the set of 1 km grid locations best represented by the environmental drivers at that particular LTAR site. Representativeness shows how well the combination of characteristics at each CONUS location was represented by the LTAR sites’ environments, while constituency shows which LTAR site was the closest match for each location. LTAR representativeness was good across most of the CONUS. Representativeness for croplands was higher than for grazinglands, probably because croplands have more specific environmental criteria. Constituencies resemble ecoregions but have their environmental conditions “centered” on those at particular existing LTAR sites. Constituency of LTAR sites can be used to prioritize the locations of experimental research at or even within particular sites, or to identify the extents that can likely be included when generalizing knowledge across larger regions of the CONUS. Sites with a large constituency have generalist environments, while those with smaller constituency areas have more specialized environmental combinations. These “specialist” sites are the best representatives for smaller, more unusual areas. The potential of sharing complementary sites from the Long-Term Ecological Research (LTER) Network and the National Ecological Observatory Network (NEON) to boost representativeness was also explored. LTAR network representativeness would benefit from borrowing several NEON sites and the Sevilleta LTER site. Later network additions must include such specialist sites that are targeted to represent unique missing environments. While this analysis exhaustively considered principal environmental characteristics related to production on working lands, we did not consider the focal agronomic systems under study, or their socio-economic context.

54 ENVIRONMENTAL SCIENCES↗

Gauges, loops, and polynomials for partition functions of graphical models

Graphical models represent multivariate and generally not normalized probability distributions. Computing the normalization factor, called the partition function, is the main inference challenge relevant to multiple statistical and optimization applications. The problem is #P-hard that is of an exponential complexity with respect to the number of variables. Here, aimed at approximating the partition function, we consider multi-graph models where binary variables and multivariable factors are associated with edges and nodes, respectively, of an undirected multi-graph. We suggest a new methodology for analysis and computations that combines the Gauge function technique from Chertkov and Chernyak with the technique developed in Anari and Oveis Gharan 2017 arXiv:1702.02937; Gurvits 2011 arXiv:1106.2844; Straszak and Vishnoi 2017 55th Annual Allerton Conf. on Communication, Control, and Computing, based on the recent progress in the field of real stable polynomials. We show that the Gauge function, representing a single-out term in a finite sum expression for the partition function which achieves extremum at the so-called belief-propagation gauge, has a natural polynomial representation in terms of gauges/variables associated with edges of the multi-graph. Moreover, Gauge function can be used to recover the partition function through a sequence of transformations allowing appealing algebraic and graphical interpretations. Algebraically, one step in the sequence consists of the application of a differential operator over gauges associated with an edge. Graphically, the sequence is interpreted as a repetitive elimination/contraction of edges resulting in multi-graph models on decreasing in size (number of edges) graphs with the same partition function as in the original multi-graph model. Even though the complexity of computing factors in the sequence of the derived multi-graph models and respective Gauge functions grow exponentially with the number of eliminated edges, polynomials associated with the new factors remain bi-stable if the original factors have this property. Moreover, we show that BP estimations in the sequence do not decrease, each low-bounding the partition function.

97 MATHEMATICS AND COMPUTING↗

Missing-Data Nonparametric Coherency Estimation

Chave recently proposed an estimator for multitaper spectral density where the time series contains missing values. In this article we generalize this technique to a multitaper estimator of coherence and phase and show that one can also obtain bootstrapped confidence intervals. Additionally, we give two examples. The first is a toy example in which the true coherence is known. In the second example we show that the multitaper missing-data coherence estimator computed on real data with a single gap comprising 11% of the data outperforms the Daniell-smoothed coherence estimator where there are no gaps. The case where the two time series have different missing indices is also discussed.

42 ENGINEERING↗

Responsiveness of miscanthus and switchgrass yields to stand age and nitrogen fertilization: A meta‐regression analysis

Abstract Optimal management of the perennial bioenergy crops, miscanthus and switchgrass, requires an understanding of their responsiveness to nitrogen (N) fertilizer at different maturity stages across locations and growing conditions. Earlier studies that have examined the yield response of these crops to N and stand age using field experiments or meta‐analysis techniques provide mixed evidence. We extend earlier studies by applying a multi‐level mixed‐effects (MLME) meta‐regression model to conduct a more extensive multivariate regression of yield response of these crops to N and stand age, while controlling for climate and location conditions and unobserved factors related to study design. Our findings are based on 1403 and 2811 yield observations for miscanthus and switchgrass, respectively, from experiments conducted between 2002 and 2019 across the rainfed region in the United States. We find statistically significant evidence that an additional year of maturity increases miscanthus and switchgrass yields but at a decreasing rate; yields peak at the 7th and 6th year respectively, for the observed range of applied N rates and stands. We also find that an increase in N application increases yield by a statistically significant level, but at a declining rate; the magnitude of the yield response to N is, however, small and varies with the age of the crop. The impact of N is larger on older compared to younger and middle‐aged stands of miscanthus. In contrast, the impact of N on switchgrass is larger on middle‐aged compared to younger and older stands of switchgrass. We do not find a statistically significant effect of soil productivity on yield for either crop. This analysis provides a basis for developing N application recommendations and optimal rotation age for miscanthus and switchgrass and shows that these energy crops can grow just as productively on low productivity land as on high productivity land.

59 BASIC BIOLOGICAL SCIENCES↗

Role of Diurnal Cycle of Insolation on the MJO Propagation in the Maritime Continent

The diurnal cycle of convection in the Maritime Continent (MC) has been hypothesized to act as a barrier to the eastward propagation of the Madden‐Julian oscillation (MJO). To test this hypothesis, we use a regional model with realistic MJO to simulate an event from the boreal spring of 2013 that weakened and stalled over the MC. Two simulations are conducted: one that includes the diurnal cycle of insolation (CTL), and another without it (NO_DC). The MJO in the simulations was identified and tracked using a large‐scale precipitation tracking method that distinguishes propagation and non‐propagation unlike the usual Real‐time Multivariate MJO method. In the NO_DC simulation, the absence of diurnal heating reduces land precipitation, allowing more continuous eastward MJO propagation. An analysis of moist static energy budget reveals that MJO maintenance in NO_DC is due to increased longwave heating and reduced advection, whereas the persistent MJO propagation in NO_DC is due to increased advection and reduced longwave heating and surface latent heat flux. These processes, however, may vary across different parts of the MC, emphasizing the complexity of MJO propagation across the MC.

Zhou, Xin [National Center for Atmospheric Researc↗

Search for ${\text {Z}{}{}} {\text {Z}{}{}} $ and ${\text {Z}{}{}} {\text {H}{}{}} $ production in the ${\text {b}{}{}} {\bar{{\text {b}{}{}}}{}{}} {\text {b}{}{}} {\bar{{\text {b}{}{}}}{}{}} $ final state using proton-proton collisions at $\sqrt{s}=13\,\text {Te}\hspace{-.08em}\text {V} $

A search for ${\text {Z}{}{}} {\text {Z}{}{}} $ and ${\text {Z}{}{}} {\text {H}{}{}} $ production in the ${\text {b}{}{}} {\bar{{\text {b}{}{}}}{}{}} {\text {b}{}{}} {\bar{{\text {b}{}{}}}{}{}} $ final state is presented, where H is the standard model (SM) Higgs boson. The search uses an event sample of proton-proton collisions corresponding to an integrated luminosity of 133$\,\text {fb}^{-1}$ collected at a center-of-mass energy of 13$\,\text {Te}\hspace{-.08em}\text {V}$ with the CMS detector at the CERN LHC. The analysis introduces several novel techniques for deriving and validating a multi-dimensional background model based on control samples in data. A multiclass multivariate classifier customized for the ${\text {b}{}{}} {\bar{{\text {b}{}{}}}{}{}} {\text {b}{}{}} {\bar{{\text {b}{}{}}}{}{}} $ final state is developed to derive the background model and extract the signal. The data are found to be consistent, within uncertainties, with the SM predictions. The observed (expected) upper limits at 95% confidence level are found to be 3.8 (3.8) and 5.0 (2.9) times the SM prediction for the ${\text {Z}{}{}} {\text {Z}{}{}} $ and ${\text {Z}{}{}} {\text {H}{}{}} $ production cross sections, respectively.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

A machine learning approach to galaxy properties: joint redshift–stellar mass probability distributions with Random Forest

We demonstrate that highly accurate joint redshift–stellar mass probability distribution functions (PDFs) can be obtained using the Random Forest (RF) machine learning (ML) algorithm, even with few photometric bands available. As an example, we use the Dark Energy Survey (DES), combined with the COSMOS2015 catalogue for redshifts and stellar masses. We build two ML models: one containing deep photometry in the griz bands, and the second reflecting the photometric scatter present in the main DES survey, with carefully constructed representative training data in each case. We validate our joint PDFs for 10 699 test galaxies by utilizing the copula probability integral transform and the Kendall distribution function, and their univariate counterparts to validate the marginals. Benchmarked against a basic set-up of the template-fitting code bagpipes, our ML-based method outperforms template fitting on all of our predefined performance metrics. In addition to accuracy, the RF is extremely fast, able to compute joint PDFs for a million galaxies in just under 6 min with consumer computer hardware. Such speed enables PDFs to be derived in real time within analysis codes, solving potential storage issues. As part of this work we have developed galpro 1, a highly intuitive and efficient python package to rapidly generate multivariate PDFs on-the-fly. galpro is documented and available for researchers to use in their cosmology and galaxy evolution studies.

79 ASTRONOMY AND ASTROPHYSICS↗

Evaluating the location capabilities of a regional infrasonic network in Utah, US, using both ray tracing-derived and empirical-derived celerity-range and backazimuth models

SUMMARY More realistic models for infrasound signal propagation across a region can be used to improve the precision and accuracy of spatial and temporal source localization estimates. Motivated by incomplete infrasound event bulletins in the Western US, the location capabilities of a regional infrasonic network of stations located between 84–458 km from the Utah Test and Training Range, Utah, USA, is assessed using a series of near-surface explosive events with complementary ground truth (GT) information. Signal arrival times and backazimuth estimates are determined with an automatic F-statistic based signal detector and manually refined by an analyst. This study represents the first application of three distinct celerity-range and backazimuth models to an extensive suite of realistic signal detections for event location purposes. A singular celerity and backazimuth deviation model was previously constructed using ray tracing analysis based on an extensive archive of historical atmospheric specifications and is applied within this study to test location capabilities. Similarly, a set of multivariate, season and location specific models for celerity and backazimuth are compared to an empirical model that depends on the observations across the infrasound network and the GT events, which accounts for atmospheric propagation variations from source to receiver. Discrepancies between observed and predicted signal celerities result in locations with poor accuracy. Application of the empirical model improves both spatial localization precision and accuracy; all but one location estimates retain the true GT location within the 90 per cent confidence bounds. Average mislocation of the events is 15.49 km and average 90 per cent error ellipse areas are 4141 km2. The empirical model additionally reduces origin time residuals; origin time residuals from the other location models are in excess of 160 s while residuals produced with the empirical model are within 30 s of the true origin time. We demonstrate that event location accuracy is driven by a combination of signal propagation model and the azimuthal gap of detecting stations. A direct relationship between mislocation, error ellipse area and increased station azimuthal gaps indicate that for sparse networks, detection backazimuths may drive location biases over traveltime estimates.

58 GEOSCIENCES↗

Evaluating the potential of short-term instrument deployment to improve distributed wind resource assessment

Distributed wind projects, which are connected at the distribution level of an electricity system or in off-grid applications to serve specific or local energy needs, often rely solely on wind resource models to establish wind speed and energy generation expectations. Historically, anemometer loan programs have provided an affordable avenue for more accurate onsite wind resource assessment, and the lowering cost of lidar systems has shown similar advantages for more recent assessments. While a full 12 months of onsite wind measurement is the standard for correcting model-based long-term wind speed estimates for utility-scale wind farms, the time and capital investment involved in gathering onsite measurements must be reconciled with the energy needs and funding opportunities that drive expedient deployment of distributed wind projects. Much literature exists to quantify the performance of correcting long-term wind speed estimates with 1 or more years of observational data, but few studies explore the impacts of correcting with months-long observational periods. This study aims to answer the question of how short you can go in terms of the observational time period needed to make impactful improvements to model-based long-term wind speed estimates. Three algorithms, multivariable linear regression, adaptive regression splines, and regression trees, are evaluated for their skill at correcting long-term wind resource estimates from the European Centre for Medium-Range Weather Forecasts Reanalysis version 5 (ERA5) using months-long periods of observational data from 66 locations across the US. On average, correction with even 1 month of observations provides significant improvement over the baseline ERA5 wind speed estimates and produces median bias magnitudes and relative errors within 0.22 m s −1 and 4 percentage points of the median bias magnitudes and relative errors achieved using the standard 12 months of data for correction. However, in cases when the shortest observational periods (1 to 2 months) used for correction are not well correlated with the overlapping ERA5 reference, the resultant long-term wind speed errors are worse than those produced using ERA5 without correction. Summer months, which are characterized by weaker relative wind speeds and standard deviations for most of the evaluation sites, tend to produce the worst results for long-term correction using months-long observations. The three tested algorithms perform similarly for long-term wind speed bias; however, regression trees perform notably worse than multivariable linear regression and adaptive regression splines in terms of correlation when using 6 months or less of observational data for correction. Translating the analysis to wind energy, median relative errors in the capacity factor are on average within 10 % using 1 month of training. If the observation period used for correction is not well correlated with the reference data, however, misrepresentation of the observed capacity factor can be substantial. The risk associated with poor correlation between the observed and reference datasets decreases with increasing training period length. In the worst-correlation scenarios, the median capacity factor relative errors from using 1, 3, and 6 months are within 47 %, 26 %, and 16 %, respectively.

17 WIND ENERGY↗

Overview of Algorithms for Using Particle Morphology in Pre-Detonation Nuclear Forensics

A major goal in pre-detonation nuclear forensics is to infer the processing conditions and/or facility type that produced radiological material. This review paper focuses on analyses of particle size, shape, texture (“morphology”) signatures that could provide information on the provenance of interdicted materials. For example, uranium ore concentrates (UOC or yellowcake) include ammonium diuranate (ADU), ammonium uranyl carbonate (AUC), sodium diuranate (SDU), magnesium diuranate (MDU), and others, each prepared using different salts to precipitate U from solution. Once precipitated, UOCs are often dried and calcined to remove adsorbed water. The products can be allowed to react further, forming uranium oxides UO3, U3O8, or UO2 powders, whose surface morphology can be indicative of precipitation and/or calcination conditions used in their production. This review paper describes statistical issues and approaches in using quantitative analyses of measurements such as particle size and shape to infer production conditions. Statistical topics include multivariate t tests (Hotelling’s T 2 ), design of experiments, and several machine learning (ML) options including decision trees, learning vector quantization neural networks, mixture discriminant analysis, and approximate Bayesian computation (ABC). ABC is emphasized as an attractive option to include the effects of model uncertainty in the selected and fitted forward model used for inferring processing conditions.

98 NUCLEAR DISARMAMENT, SAFEGUARDS, AND PHYSICAL P↗

Transformative Impacts of Laser-Induced Breakdown Spectroscopy on Environmental and Biological Research at Oak Ridge National Laboratory

This manuscript will present an advancement of transformative research that has been conducted at Oak Ridge National Laboratory (ORNL) over a 25-year period (2000–2025) on a variety of environmental and biological matrices. These investigations derived a fundamental understanding of how elemental detection and analysis of these matrices led to the knowledge and discovery of natural processes in plants and the environment. Each project led to the initiation of a new research area which unearthed awesome and novel breakthroughs. Highlights are listed below: 1. The preliminary research at ORNL centered on the detection of aerosols utilizing Laser-induced Breakdown Spectroscopy (LIBS) technology. The Clean Air Act Amendment (CAAA) of 1990 highlighted the importance of identifying hazardous air pollutants (HAPs) due to their impact on environmental and human health, thereby underscoring the need to detect various toxic elements. Research in aerosol chemistry aimed to identify these harmful elements released by factories during periods of increased emissions in their manufacturing processes. LIBS emerged as the most effective method for real-time, in situ measurements of metal species in both gaseous and aerosol phases. 2. An understanding of the presence of total carbon in soils gives perspective on how to develop carbon sequestration strategies. The recognition that carbon sinks can evolve back to carbon sources to emit back to the atmosphere was an important consideration. Also, the concentration of carbon in soil indicates the health of land areas for growing crops successfully. 3. The direct detection of most of the elements in a wood sample in a single emission spectrum, without sample preparation, encouraged the research to use the LIBS technique for preservative treated wood coupled with use of multivariate statistical methodology. Additionally, it encouraged the researchers to try to differentiate natural woods from different parts of the country, and it was successfully demonstrated that LIBS coupled with MVA analysis could differentiate wood of different species from each other and of similar species grown in different environments based on their elemental spectra. This was a breakthrough since it revealed a systematic approach to connect elemental scarcity and abundance to either drought or typical rainfall conditions for the hardwood trees grown in specific areas. 4. Furthermore, the research progressed to reveal physiological and developmental processes contributing to biomass production such that the variation in leaf elemental composition increases our understanding of terrestrial nutrient cycles, as well as tracking the transfer of toxic elements from soils to living organisms. 5. Recently another breakthrough viz., ionomics initiated the correlation of elements to specific genes, uncovering the function that the element performed in the plant. More recently, this has been extended from plants to fungi as well as fungi growing in symbiotic relations with plants.

09 BIOMASS FUELS↗

A Novel Non‐Destructive Rapid Tool for Estimating Amino Acid Composition and Secondary Structures of Proteins in Solution

Abstract Amino‐acid protein composition plays an important role in biology, medicine, and nutrition. Here, a groundbreaking protein analysis technique that quickly estimates amino acid composition and secondary structure across various protein sizes, while maintaining their natural states is introduced and validated. This method combines multivariate statistics and the thermostable Raman interaction profiling (TRIP) technique, eliminating the need for complex preparations. In order to validate the approach, the Raman spectra are constructed of seven proteins of varying sizes by utilizing their amino acid frequencies and the Raman spectra of individual amino acids. These constructed spectra exhibit a close resemblance to the actual measured Raman spectra. Specific vibrational modes tied to free amino and carboxyl termini of the amino acids disappear as signals linked to secondary structures emerged under TRIP conditions. Furthermore, the technique is used inversely to successfully estimate amino acid compositions and secondary structures of unknown proteins across a range of sizes, achieving impressive accuracy ranging between 1.47% and 5.77% of root mean square errors (RMSE). These results extend the uses for TRIP beyond interaction profiling, to probe amino acid composition and structure.

Chemistry↗

The genetic basis for panicle trait variation in switchgrass ( Panicum virgatum )

Grass species exhibit large diversity in panicle architecture influenced by genes, the environment, and their interaction. The genetic study of panicle architecture in perennial grasses is limited. In this study, we evaluate the genetic basis of panicle architecture including panicle length, primary branching number, and secondary branching number in an outcrossed switchgrass QTL population grown across ten field sites in the central USA through multi-environment mixed QTL analysis. We also evaluate genetic effects in a diversity panel of switchgrass grown at three of the ten field sites using genome-wide association (GWAS) and multivariate adaptive shrinkage. Furthermore, we search for candidate genes underlying panicle traits in both of these independent mapping populations. Overall, 18 QTL were detected in the QTL mapping population for the three panicle traits, and 146 unlinked genomic regions in the diversity panel affected one or more panicle trait. Twelve of the QTL exhibited consistent effects (i.e., no QTL by environment interactions or no QTL × E), and most (four of six) of the effects with QTL × E exhibited site-specific effects. Most (59.3%) significant partially linked diversity panel SNPs had significant effects in all panicle traits and all field sites and showed pervasive pleiotropy and limited environment interactions. Panicle QTL co-localized with significant SNPs found using GWAS, providing additional power to distinguish between true and false associations in the diversity panel.

59 BASIC BIOLOGICAL SCIENCES↗

Advanced Method Optimization for Sampling and Analysis Instrumentation

This work presents a generalized approach for analytical method optimization that branches the gap between techniques historically employed and accurate modern optimization techniques suitable for various applications. The novelty of the described strategy is the utilization of multivariate, multiobjective optimization with Karush-Kuhn-Tucker conditions to bound the optimization space to solutions within the physical limitations of instrumentation. Briefly, the basic steps outlined in this paper are to (1) determine the objective(s) that should be maximized or minimized based on the goals of the analytical application, (2) conduct a screening experiment, (3) perform ANOVA to determine the parameters which have a statistically significant effect on the objective, (4) conduct an experiment (e.g., Box-Behnken design) to collect data for fitting the objective equation, and (5) determine the physical constraints of the parameters and solve the Lagrangian to determine the optimal method parameters. A broad approach to optimization target selection allows for robust method tuning to develop improved data sets amenable for chemometrics and machine learning algorithm development. Gas chromatography-mass spectrometry was selected as a use case due to its broad use across scientific fields and time-consuming method development involving numerous parameters. In conclusion, this strategy can reduce the cost of research, improve data quality, and enable the rapid development of new analytical technique.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗