Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Consensus”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 127 records · Page 7

Pseudouridine synthase 7 is an opportunistic enzyme that binds and modifies substrates with diverse sequences and structures

Pseudouridine (Ψ) is a ubiquitous RNA modification incorporated by pseudouridine synthase (Pus) enzymes into hundreds of noncoding and protein-coding RNA substrates. Here, in this study, we determined the contributions of substrate structure and protein sequence to binding and catalysis by pseudouridine synthase 7 (Pus7), one of the principal messenger RNA (mRNA) modifying enzymes. Pus7 is distinct among the eukaryotic Pus proteins because it modifies a wider variety of substrates and shares limited homology with other Pus family members. We solved the crystal structure of Saccharomyces cerevisiae Pus7, detailing the architecture of the eukaryotic-specific insertions thought to be responsible for the expanded substrate scope of Pus7. Additionally, we identified an insertion domain in the protein that fine-tunes Pus7 activity both in vitro and in cells. These data demonstrate that Pus7 preferentially binds substrates possessing the previously identified UGUAR (R = purine) consensus sequence and that RNA secondary structure is not a strong requirement for Pus7-binding. In contrast, the rate constants and extent of Ψ incorporation are more influenced by RNA structure, with Pus7 modifying UGUAR sequences in less-structured contexts more efficiently both in vitro and in cells. Although less-structured substrates were preferred, Pus7 fully modified every transfer RNA, mRNA, and nonnatural RNA containing the consensus recognition sequence that we tested. Our findings suggest that Pus7 is a promiscuous enzyme and lead us to propose that factors beyond inherent enzyme properties (e.g., enzyme localization, RNA structure, and competition with other RNA-binding proteins) largely dictate Pus7 substrate selection.

59 BASIC BIOLOGICAL SCIENCES↗

CONSTAX2: improved taxonomic classification of environmental DNA markers

Abstract Summary CONSTAX—the CONSensus TAXonomy classifier—was developed for accurate and reproducible taxonomic annotation of fungal rDNA amplicon sequences and is based upon a consensus approach of RDP, SINTAX and UTAX algorithms. CONSTAX2 extends these features to classify prokaryotes as well as eukaryotes and incorporates BLAST-based classifiers to reduce classification errors. Additionally, CONSTAX2 implements a conda-installable command-line tool with improved classification metrics, faster training, multithreading support, capacity to incorporate external taxonomic databases and new isolate matching and high-level taxonomy tools, replete with documentation and example tutorials. Availability and implementation CONSTAX2 is available at https://github.com/liberjul/CONSTAXv2, and is packaged for Linux and MacOS from Bioconda with use under the MIT License. A tutorial and documentation are available at https://constax.readthedocs.io/en/latest/. Data and scripts associated with the manuscript are available at https://github.com/liberjul/CONSTAXv2_ms_code. Supplementary information Supplementary data are available at Bioinformatics online.

59 BASIC BIOLOGICAL SCIENCES↗

A best-practice guide to predicting plant traits from leaf-level hyperspectral data using partial least squares regression

Partial least squares regression (PLSR) modelling is a statistical technique for correlating datasets, and involves the fitting of a linear regression between two matrices. One application of PLSR enables leaf traits to be estimated from hyperspectral optical reflectance data, facilitating rapid, high-throughput, non-destructive plant phenotyping. This technique is of interest and importance in a wide range of contexts including crop breeding and ecosystem monitoring. The lack of a consensus in the literature on how to perform PLSR means that interpreting model results can be challenging, applying existing models to novel datasets can be impossible, and unknown or undisclosed assumptions can lead to incorrect or spurious predictions. We address this lack of consensus by proposing best practices for using PLSR to predict plant traits from leaf-level hyperspectral data, including a discussion of when PLSR is applicable, and recommendations for data collection. Further, we provide a tutorial to demonstrate how to develop a PLSR model, in the form of an R script accompanying this manuscript. This practical guide will assist all those interpreting and using PLSR models to predict leaf traits from spectral data, and advocates for a unified approach to using PLSR for predicting traits from spectra in the plant sciences.

54 ENVIRONMENTAL SCIENCES↗

Structures of Neisseria gonorrhoeae MtrR-operator complexes reveal molecular mechanisms of DNA recognition and antibiotic resistance-conferring clinical mutations

Abstract Mutations within the mtrR gene are commonly found amongst multidrug resistant clinical isolates of Neisseria gonorrhoeae, which has been labelled a superbug by the Centers for Disease Control and Prevention. These mutations appear to contribute to antibiotic resistance by interfering with the ability of MtrR to bind to and repress expression of its target genes, which include the mtrCDE multidrug efflux transporter genes and the rpoH oxidative stress response sigma factor gene. However, the DNA-recognition mechanism of MtrR and the consensus sequence within these operators to which MtrR binds has remained unknown. In this work, we report the crystal structures of MtrR bound to the mtrCDE and rpoH operators, which reveal a conserved, but degenerate, DNA consensus binding site 5′-MCRTRCRN4YGYAYGK-3′. We complement our structural data with a comprehensive mutational analysis of key MtrR-DNA contacts to reveal their importance for MtrR-DNA binding both in vitro and in vivo. Furthermore, we model and generate common clinical mutations of MtrR to provide plausible biochemical explanations for the contribution of these mutations to multidrug resistance in N. gonorrhoeae. Collectively, our findings unveil key biological mechanisms underlying the global stress responses of N. gonorrhoeae.

59 BASIC BIOLOGICAL SCIENCES↗

Benefits and Limits of Phasing Alleles for Network Inference of Allopolyploid Complexes

Abstract Accurately reconstructing the reticulate histories of polyploids remains a central challenge for understanding plant evolution. Although phylogenetic networks can provide insights into relationships among polyploid lineages, inferring networks may be hindered by the complexities of homology determination in polyploid taxa. We use simulations to show that phasing alleles from allopolyploid individuals can improve phylogenetic network inference under the multispecies coalescent by obtaining the true network with fewer loci compared with haplotype consensus sequences or sequences with heterozygous bases represented as ambiguity codes. Phased allelic data can also improve divergence time estimates for networks, which is helpful for evaluating allopolyploid speciation hypotheses and proposing mechanisms of speciation. To achieve these outcomes in empirical data, we present a novel pipeline that leverages a recently developed phasing algorithm to reliably phase alleles from polyploids. This pipeline is especially appropriate for target enrichment data, where the depth of coverage is typically high enough to phase entire loci. We provide an empirical example in the North American Dryopteris fern complex that demonstrates insights from phased data as well as the challenges of network inference. We establish that our pipeline (PATÉ: Phased Alleles from Target Enrichment data) is capable of recovering a high proportion of phased loci from both diploids and polyploids. These data may improve network estimates compared with using haplotype consensus assemblies by accurately inferring the direction of gene flow, but statistical nonidentifiability of phylogenetic networks poses a barrier to inferring the evolutionary history of reticulate complexes.

Evolutionary Biology↗

Matrix Metalloproteinases as Candidate Antigenic Determinants for Anti‐Tumor Autoantibodies in Human Ovarian Cancer: A Post Hoc Analysis

Circulating antibodies in patients with cancer can facilitate the identification of accessible epitopes on autoantigens expressed by tumors. To identify previously unrecognized protein targets in ovarian cancer, we computationally assessed a heptapeptide consensus motif (VPELGHE, flanked by two cysteine residues yielding a cyclic nonapeptide under oxidizing conditions) previously discovered via phage display-based epitope mapping of autoantibodies in patients. Eight proteins associated with ovarian cancer encompass amino acid sequences similar to the consensus motif and were, therefore, considered as candidate native autoantigens. Among these candidate targets, however, matrix metalloproteinase 14 (MMP14) demonstrates gene expression that is both high and negatively correlated with survival in ovarian cancer patient cohorts. MMP14 protein levels are also stable in tumor versus non-tumor tissues. Moreover, the corresponding heptapeptide mimic in MMP14 occurs within an α-helical secondary structural element observed in its catalytic domain. These findings demonstrate that a subset of patient-derived autoantibodies may interact with a previously unknown antigenic epitope found in MMP14 and other MMPs, thereby providing opportunities for the development of new targeted agents.

Biochemistry & Molecular Biology↗

Differential analysis of incompressibility in neutron-rich nuclei

Both the incompressibility K A of a finite nucleus of mass A and that (K ∞ ) of infinite nuclear matter are fundamentally important for many critical issues in nuclear physics and astrophysics. While some consensus has been reached about K ∞ , accurate theoretical predictions and experimental extractions of K τ characterizing the isospin dependence of K A have been very difficult. We propose a differential approach to extract K τ and K ∞ independently from the K A data of any two nuclei in a given isotope chain. Applying this method to the K A data from isoscalar giant monopole resonances (ISGMR) in even-even Pb, Sn, Cd, and Ca isotopes taken by Garg et al. at the Research Center for Nuclear Physics (RCNP), Osaka University, Japan, we find that the 106 Cd– 116 Cd and 112 Sn– 124 Sn pairs having the largest differences in isospin asymmetries in their respective isotope chains measured so far provide consistently the most accurate up-to-date K τ value of K τ = –616 ± 59 MeV and K τ =–623 ± 86 MeV, respectively, largely independent of the remaining uncertainties of the surface and Coulomb terms in expanding K A , while the K ∞ values extracted from different isotopes chains are all well within the current uncertainty range of the community consensus for K ∞ . Moreover, the size and origin of the “soft Sn puzzle” is studied with respect to the “stiff Pb phenomenon.” Furthermore, it is found that the latter is favored due to a much larger (by ≈ 380 MeV) K τ for Pb isotopes than for Sn isotopes, while K ∞ from analyzing the K A data of Sn isotopes is only about 5 MeV less than that from analyzing the Pb data.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Distributed ADMM Using Private Blockchain for Power Flow Optimization in Distribution Network With Coupled and Mixed-Integer Constraints

The optimization problem for scheduling distributed energy resources (DERs) and battery energy storage systems (BESS) integrated with the power grid is important to minimize energy consumption from conventional sources in response to demand. Conventionally this optimization problem is solved in a centralized manner, limiting the size of the problem that can be solved and creating a high communication overhead because all the data is transferred to the central controller. These limitations are addressed by the proposed distributed consensus-based alternating direction method of multiplier (DC-ADMM) optimization algorithm, which decomposes the optimization problem into subproblems with private cost function and constraints. The distribution feeder is partitioned into low coupling subnetworks/regions, which solves the private subproblem locally and exchanges information with the neighboring regions to reach consensus. The relaxation strategy is employed for mixed-integer and coupled constraints introduced in the optimal power flow (OPF) problem by stationary and transportable BESS because DC-ADMM convergence is only guaranteed for strict convex problems. The information exchange and synchronization between subnetworks/regions are vital for distributed optimization. In this work, both of these aspects are addressed by the blockchain. The smart contract deployed on the blockchain network acts as a mediator for secure data exchange and synchronization in distributed computation. The blockchain-based distributed optimization problem’s effectiveness is tested for a 0.5-MW laboratory microgrid for one hour ahead and day-ahead for the IEEE 123-bus and EPRI J1 test feeders, and results are compared with a centralized solution.

25 ENERGY STORAGE↗

Game Theoretic Orchestration for Cooperation among Power Distribution System Applications

The evolving transformation with the proliferation of distributed energy resources and advanced metering, necessitates advanced distribution systems to integrate and orchestrate a large number of grid-edge devices while also serving multiple system-level objectives such as resilience, decarbonization, equity and other system mandates. The parallel deployment and control of resources towards achieving diverse objectives may lead to conflicts between applications that want to control overlapping sets of device setpoints, potentially leading to oscillatory behavior and suboptimal performance. This work aims at leveraging game theoretic framework to drive cooperative behavior among competitive applications. The work proposes a weighted-consensus based game design to facilitate conflict resolution through consensus-building iterations for modular platform. Simulation-based evaluation on a sample test system demonstrates the performance the proposed deconfliction strategy in resolving operational conflicts and achieving close-to-optimal trade off among the applications. Results also compare the proposed strategy with a distribution optimization approach and illustrate it effectiveness in diverse apps regardless of their design while also incentivizing apps with flexible design.

Advanced distribution operations, cooperation, app↗

Decentralized Collaborative Learning with Probabilistic Data Protection

We discuss future directions of Blockchain as a collaborative value co-creation platform, in which network participants can gain extra insights that cannot be accessed when disconnected from the others. As such, we propose a decentralized machine learning framework that is carefully designed to respect the values of democracy, diversity, and privacy. Specifically, we propose a federated multi-task learning framework that integrates a privacy-preserving dynamic consensus algorithm. We show that a specific network topology called the expander graph dramatically improves the scalability of global consensus building. We conclude the paper by making some remarks on open problems.

Ide, Tsuyoshi↗

A Unified Testing Platform to Mature Blockchain Applications for Grid Emulation Environments

Blockchain technology is a relatively novel technology that can be used to develop more decentralized, autonomous and tamper-evident solutions. A feature that can aid Transactive Energy Systems to reach their goals by enabling individual actors to communicate and reach consensus with other participants in a more decentralized fashion. However, technical barriers to evaluate and adopt this type of technology within the electrical domain still exist. To facilitate this task, BLOSEM Unified Testing Platform (UTP), a DOE-sponsored, multi-lab effort intends to accelerate the development of solutions by offering a common set of reusable services that can be used to interconnect existent grid tools with blockchain services. UTP is intended to serve as development platform that can provide application engineers with the technical means to evaluate potential blockchain solutions, by enabling them to concentrate on the actual application functionalities while at the same time abstracting the connectivity and performance measurement tasks. The use of BLOSEM UTP is further demonstrated by implementing two potential use cases that are intended to validate both the feasibility of implementing these applications as blockchain-based solutions while also demonstrating the features provided by UTP.

blockchain co-simulation↗

Fast Tuning-Free Distributed Algorithm for Solving the Network-Constrained Economic Dispatch

With the increasing penetration of distributed energy resources (DERs) and their participation in the electricity market, it becomes more desirable to apply distributed algorithms for resource allocation in order to address the resulting computational and communicational challenges. Most of the existing distributed algorithms for solving the network-constrained economic dispatch (NCED) problem require the tuning of certain auxiliary parameters. As a result, the robustness of these algorithms against the varieties in DERs is greatly undermined. In this paper, a new distributed algorithm, optimality condition consensus (OCC), is proposed to solve the NCED problem by using distributed power flow (DPF) and ratio consensus as fundamental tools. It inherits the advantages of existing distributed algorithms for the NCED problem but removes the need for parameter tuning to improve performance in practice. In conclusion, the effectiveness of the proposed distributed algorithm in terms of efficiency, scalability, and robustness is demonstrated through detailed case studies.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Multi-Model Future Typical Meteorological (fTMY) Weather Files for nearly every US County

Exploring climate-induced impacts on building energy consumption can provide valuable insights for sustainable energy planning and environmental management in the face of a changing climate. By utilizing future weather data statistically downscaled from the Intergovernmental Panel on Climate Change (IPCC) General Circulation Models (GCMs) from 2020-2100, this paper presents a broadening industry-consensus approach for generating future Typical Meteorological Year (fTMY) weather files through a combination of statistical downscaling and high-performance computing that generalizes across decades, multiple locations for a region, and varying climate models. Furthermore, these fTMY files have been generated for 3,128 US counties for capturing potential weather on a 20-year basis.

Building↗

Blockchain Applicability and Cybersecurity Frameworks (BAF)

The Blockchain Applicability and Cybersecurity Frameworks is the implementation of an earlier invention disclosure and Provisional Patent Application: S-146,805, Battelle IPID 31608-E. —BAF walks a user (or an interested party) through ~100 controls to determine various factors, including: 1. Does the application need a blockchain? 2. If the application needs a blockchain, does it need a private blockchain or permissionless/public blockchain? 3. What kind of consensus is most suitable for the application? The current version of BAF evaluates between four consensus mechanisms: proof-of-work, proof-of-stake, proof-of-burn, and proof-of-authority.

Gourisetti, Sri Nikhil Gupta↗

Microbiome data management in action workshop: Atlanta, GA, USA, June 12–13, 2024

Microbiome research is revolutionizing human and environmental health, but the value and reuse of microbiome data are significantly hampered by the limited development and adoption of data standards. While several ongoing efforts are aimed at improving microbiome data management, significant gaps still remain in terms of defining and promoting adoption of consensus standards for these datasets. The Strengthening the Organization and Reporting of Microbiome Studies (STORMS) guidelines for human microbiome research have been endorsed and successfully utilized by many research organizations, publishers, and funding agencies, and have been recognized as a consensus community standard. No equivalent effort has occurred for environmental, synthetic, and non-human host-associated microbiomes. To address this growing need within the microbiome research community, we convened the Microbiome Data Management in Action Workshop (June 12–13, 2024, in Atlanta, GA, USA), to bring together key decision makers in microbiome science including researchers, publishers, funders, and data repositories. The 50 attendees, representing the diverse and interdisciplinary nature of microbiome research, discussed recent progress and challenges, and brainstormed actionable recommendations and paths forward for coordinated environmental microbiome data management and the modifications necessary for the STORMS guidelines to be applied to environmental, non-human host, and synthetic microbiomes. The outcomes of this workshop will form the basis of a formalized data management roadmap to be implemented across the field. These best practices will drive scientific innovation now and in years to come as these data continue to be used not only in targeted reanalyses but in large-scale models and machine learning efforts.

54 ENVIRONMENTAL SCIENCES↗

CATMoS: Collaborative Acute Toxicity Modeling Suite

Background: Humans are exposed to tens of thousands of chemical substances that need to be assessed for their potential toxicity. Acute systemic toxicity testing serves as the basis for regulatory hazard classification, labeling, and risk management. However, it is cost- and time-prohibitive to evaluate all new and existing chemicals using traditional rodent acute toxicity tests. In silico models built using existing data facilitate rapid acute toxicity predictions without using animals. Objectives: The U.S. Interagency Coordinating Committee on the Validation of Alternative Methods Acute Toxicity Workgroup organized an international collaboration to develop in silico models for predicting acute oral toxicity based on five different endpoints: LD50 value, U.S. Environmental Protection Agency hazard categories, Globally Harmonized System for Classification and Labelling hazard categories, very toxic chemicals (LD50 =50 mg/kg), and non-toxic chemicals (LD50 >2000 mg/kg). Methods: An acute oral toxicity data inventory for 11,992 chemicals was compiled, split into training and evaluation sets, and made available to 35 participating international research groups that submitted a total of 139 predictive models. Predictions that fell within the applicability domains of the submitted models were evaluated using external validation sets. These were then combined into consensus models to leverage strengths of individual approaches. Results: The resulting consensus predictions, which leverage the collective strengths of each individual model, form the Collaborative Acute Toxicity Modeling Suite (CATMoS). CATMoS demonstrated high performance in terms of accuracy and robustness when compared to in vivo results. Discussion: CATMoS is being evaluated by regulatory agencies for its utility and applicability as a potential replacement for in vivo rat acute oral toxicity studies. CATMoS predictions for over 800,000 chemicals have been made available via the NTP’s Integrated Chemical Environment. The models are also implemented in a free, standalone open-source tool, OPERA, which allows predictions of new and untested chemicals to be made.

63 RADIATION, THERMAL, AND OTHER ENVIRON. POLLUTAN↗

DEEPEN 3D PFA Weights for Exploration Datasets in Magmatic Environments

DEEPEN stands for DE-risking Exploration of geothermal Plays in magmatic ENvironments. As part of the development of the DEEPEN 3D play fairway analysis (PFA) methodology for magmatic plays (conventional hydrothermal, superhot EGS, and supercritical), weights needed to be developed for use in the weighted sum of the different favorability index models produced from geoscientific exploration datasets. This GDR submission includes those weights. The weighting was done using two different approaches: one based on expert opinions, and one based on statistical learning. The weights are intended to describe how useful a particular exploration method is for imaging each component of each play type. They may be adjusted based on the characteristics of the resource under investigation, knowledge of the quality of the dataset, or simply to reduce the impact a single dataset has on the resulting outputs. Within the DEEPEN PFA, separate sets of weights are produced for each component of each play type, since exploration methods hold different levels of importance for detecting each play component, within each play type. The weights for conventional hydrothermal systems were based on the average of the normalized weights used in the DOE-funded PFA projects that were focused on magmatic plays. This decision was made because conventional hydrothermal plays are already well-studied and understood, and therefore it is logical to use existing weights where possible. In contrast, a true PFA has never been applied to superhot EGS or supercritical plays, meaning that exploration methods have never been weighted in terms of their utility in imaging the components of these plays. To produce weights for superhot EGS and supercritical plays, two different approaches were used: one based on expert opinion and the analytical hierarchy process (AHP), and another using a statistical approach based on principal component analysis (PCA). The weights are intended to provide standardized sets of weights for each play type in all magmatic geothermal systems. Two different approaches were used to investigate whether a more data-centric approach might allow new insights into the datasets, and also to analyze how different weighting approaches impact the outcomes. The expert/AHP approach involved using an online tool (https://bpmsg.com/ahp/) with built-in forms to make pairwise comparisons which are used to rank exploration methods against one-another. The inputs are then combined in a quantitative way, ultimately producing a set of consensus-based weights. To minimize the burden on each individual participant, the forms were completed in group discussions. While the group setting means that there is potential for some opinions to outweigh others, it also provides a venue for conversation to take place, in theory leading the group to a more robust consensus then what can be achieved on an individual basis. This exercise was done with two separate groups: one consisting of U.S.-based experts, and one consisting of Iceland-based experts in magmatic geothermal systems. The two sets of weights were then averaged to produce what we will from here on refer to as the "expert opinion-based weights," or "expert weights" for short. While expert opinions allow us to include more nuanced information in the weights, expert opinions are subject to human bias. Data-centric or statistical approaches help to overcome these potential human biases by focusing on and drawing conclusions from the data alone. More information on this approach along with the dataset used to produce the statistical weights may be found in the linked dataset below.

15 GEOTHERMAL ENERGY↗

FY21 Progress Report: SRNL Analysis of ICCWR LCM and WAMS data for Corrosion and Cracking

The development of algorithms for machine learning and data analysis for the 3013 Surveillance Program is a collaborative effort by the Savannah River National Laboratory (SRNL) and the University of South Carolina (USC). For corrosion detection, Laser Confocal Microscope (LCM) or Wide Area 3D Measurement System (WAMS) data is extracted from large binary files, with software written to convert the data to physical attributes (e.g., height, color and grayscale values; all as functions of a location in a plane projection). A user-friendly Matlab Graphical User Interface (GUI) that reads data from either LCM or WAMS files was developed to integrate input data with software developed for processing and evaluation. The GUI can selectively download binary data, interrogate data attributes, label data, flag significant features, execute Machine Learning (ML) algorithms, output parameters for trained ML algorithms, report ML model accuracy with respect to labeled data, and generate graphical representations for various analyses. Features can be called out by user-specified thresholds, manual labeling or machine learning algorithms when they have been completed. The ability to rapidly label data is important because of the volume of data required for training machine learning algorithms. The GUI has the flexibility to allow addition of improved ML algorithms, methods for data visualization, and statistical computations. Statistical analyses via the GUI include areas of pits within a defined range of pit depths, correlations between Red-Green-Blue (RGB) or grayscale intensity and relative surface height, covariances between values associated with features, and feature histograms. The development of supervised machine learning algorithms, however, has been hindered by a lack of training data. The machine learning algorithms for crack identification are being refined but require improvements to the true positive rate for crack detection. This shortcoming is an artifact of the limited training data currently available, perhaps more so than the structure of the neural networks. At present, the best results are had from a consensus over an ensemble of randomly generated Deep Neural Network (DNN) or Convolutional Neural Network (CNN) algorithms. Although the consensus accuracy method has yielded optimum true positive and true negative rates in excess of 80%, additional validation testing is necessary. In addition to the suite of LCM data that was initially used, and which represents the majority of the work presented in this report, WAMS image data was also reviewed at a preliminary level. The review included a comparison between image resolution and dynamic range for each method. WAMS (ZON file) image data was found to have a pixel pitch of 3.69μm compared to 1 μm for the LCM (vk4 file) data, which implies a lower resolution for the WAMS images. Conversely, the ratio of dynamic range of the WAMS data to the LCM data was approximately 41:20 for height data, suggesting that information from WAMS should more accurately determine the depth of pits. At present, the significance of the greater dynamic range of the WAMS data relative to the LCM data has not yet been evaluated.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗