Engineering PapersSearch

SEARCH · Engineering Papers

Results for “Security Genomics”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Genomics-Based Security Protocols: From Plaintext to Cipherprotein

The evolving nature of the internet will require continual advances in authentication and confidentiality protocols. Nature provides some clues as to how this can be accomplished in a distributed manner through molecular biology. Cryptography and molecular biology share certain aspects and operations that allow for a set of unified principles to be applied to problems in either venue. A concept for developing security protocols that can be instantiated at the genomics level is presented. A DNA (Deoxyribonucleic acid) inspired hash code system is presented that utilizes concepts from molecular biology. It is a keyed-Hash Message Authentication Code (HMAC) capable of being used in secure mobile Ad hoc networks. It is targeted for applications without an available public key infrastructure. Mechanics of creating the HMAC are presented as well as a prototype HMAC protocol architecture. Security concepts related to the implementation differences between electronic domain security and genomics domain security are discussed.

Shaw, Harry

Genomics and Proteomics Based Security Protocols for Secure Network Architectures

A hardware design that integrates live and algorithmic inhabitants to produce patterns of gene expression in vivo and in silico Protocols and algorithms based upon the processes of regulation of gene expression to produce cryptographic representations of genes, RNA, proteins, and gene expression to perform authentication and confidentiality functions for computers and networks. A network concept of operations integrating all of the above into existing legacy networks.

Security Genomics

IMAGINE BioSecurity: Mesocosm-Based Methods to Evaluate Biocontainment Strategies and Impact of Industrial Microbes Upon Native Ecosystems

Project Goals: The Integrative Modeling and Genome-scale Engineering for Biosystems Security (IMAGINE BioSecurity) SFA project seeks to establish an understanding of the behavior of engineered microbes in controlled versus environmental conditions to predictively devise new strategies for responding to biological escape. To this end, the IMAGINE Team has established a plant-soil mesocosm platform to track and quantify the fate of industrial microbes in environmental systems and assess the efficacy of biocontainment constraints upon genetically engineered microbe escape frequency and the impact of industrial microbes upon native ecological microbiomes. Abstract Text: Genetically modified industrial production microbes and their associated bioproducts have emerged as an integral component of a sustainable bioeconomy. However, the rapid development of these innovative technologies raises biosecurity concerns, namely, the risk of environmental escape. Thus, the realization of a bioeconomy hinges not only on the development and deployment of microbial production hosts, but also on the development of secure biosystems and biocontainment designs. Current laboratory-based biocontainment testing systems do not accurately reflect complexities found in natural environments, necessitating an environmentally relevant analysis pipeline that allows for the detection of rare escapees, the effect of associated bio-products, and the impact on native ecologies. To this end, we have developed an approach that utilizes soil mesocosms and integrated systems analyses to evaluate the efficacy of novel biocontainment strategies and to assess the impact of production systems upon terrestrial microbiome dynamics. We demonstrate the utility of this approach by modeling a contamination with industrial microbial chasses versus their biocontained counterparts. Here we demonstrate the broad utility of this system by highlighting findings from both strains of Saccharomyces cerevisiae that are contained with an inducible toxin anti-toxin system, and stains of Escherichia coli that are contained via genomic recoding. The resultant data demonstrate that this system has broad utility across diverse microbial chassis and biocontainment strategies, enables us to track the fate of our contaminating microbe with high sensitivity in the soil, as well as monitor broader impacts of the perturbation on the underlying soil system. The findings presented here support the use of this mesocosm-based approach to assess the environmental impact of industrial microbes and to validate biocontainment strategies.

BASIC BIOLOGICAL SCIENCES,INORGANIC, ORGANIC, PHYS

Methods for safely sharing dual-use genetic data

Background: Some genetic data has dual-use potential. Sharing pathogen data has shown tremendous value. For example therapeutic development and lineage tracking during the COVID pandemic. This data sharing is complicated by the fact that these data have the potential to be used for harm. The genome sequence of a pathogen can be used to enable malicious genetic engineering approaches or to recreate the pathogen from synthetic DNA. Standard data security methods can be applied to genetic data, but when data is shared between institutions, ensuring appropriate security can be difficult. Sensitive data that is shared internationally among a wide array of institutions can be especially difficult to control. Methods for securely storing and sharing genetic data with potential for dual-use are needed to mitigate this potential harm.Results: Here we propose new methods that allow genetic data to be shared in a data format that prevents a nefarious actor from accessing sensitive aspects of the data. Our methods obfuscate raw sequence data by pooling reads from different samples. This approach can ensure that data is secure while stored and during electronic transfer. We demonstrate that by pooling raw sequence data from multiple samples of the same organism, the ability to fully reconstruct any individual sample is prevented. In the pooled data, most genomic information remains, but reads or mutations cannot be directly attributed to any individual sample. To further restrict access to information, regions of a genome can be removed from the reads.Conclusion: Our methods obscure genomic information within raw sequence reads. This method can allow genetic data to be stored and shared while preventing a nefarious actor from being able to perfectly reconstruct an organism. Broad-scale sequence information remains, while fine scale details about specific samples are difficult or impossible to reconstruct. Our software is available at https://github.com/Geneinfosec-Inc/ReadMixer.

59 BASIC BIOLOGICAL SCIENCES

A Survey on the Expanding Scope and Interdisciplinary Opportunities for Processing-in-Memory Techniques

Processing-in-Memory (PIM) is emerging as a practical path to overcome the limitations of traditional von Neumann architectures. At its core, PIM systems implement computing primitives such as logic operations and multiply-accumulate acceleration through compute-in-memory, near-memory processing, or hybrid designs. The role of memory cells varies widely across technologies, acting as inputs, outputs, or analog accumulators through bit-lines and sense amplifiers. This diversity creates trade-offs in precision, bandwidth, latency, and programmability, making it difficult to build a unified understanding on the progress of the field. In this survey, we organize recent advances of PIM into three areas. First, we discuss the progress on the architectural optimizations of PIM and its integration with both DRAM and emerging non-volatile memories. Second, we examine how PIM is being used to accelerate key computing domains, including generative AI workloads and high-performance kernels, along with new approaches. Third, we highlight the growing adoption of PIM in computational sciences, where it is being applied to solve interdisciplinary problems such as genome analysis, mRNA quantification, mass spectrometry, quantum circuit simulation, wave modeling, and secure computation. Finally, we synthesize the major challenges that continue to slow PIM adoption, including manufacturing constraints, power delivery, thermal reliability, data consistency, runtime and memory-management coordination, and the difficulty of building portable software abstractions without sacrificing commercial viability. This work provides an updated, structured perspective on PIM’s potential across computing and computational sciences and the barriers that must be solved for it to reach its full impact.

Asifuzzaman, Kazi [Oak Ridge National Laboratory (

Building a FAIR data ecosystem for incorporating single-cell transcriptomics data into agricultural genome to phenome research

Introduction The agriculture genomics community has numerous data submission standards available, but the standards for describing and storing single-cell (SC, e.g., scRNA- seq) data are comparatively underdeveloped. Methods To bridge this gap, we leveraged recent advancements in human genomics infrastructure, such as the integration of the Human Cell Atlas Data Portal with Terra, a secure, scalable, open-source platform for biomedical researchers to access data, run analysis tools, and collaborate. In parallel, the Single Cell Expression Atlas at EMBL-EBI offers a comprehensive data ingestion portal for high-throughput sequencing datasets, including plants, protists, and animals (including humans). Developing data tools connecting these resources would offer significant advantages to the agricultural genomics community. The FAANG data portal at EMBL-EBI emphasizes delivering rich metadata and highly accurate and reliable annotation of farmed animals but is not computationally linked to either of these resources. Results Herein, we describe a pilot-scale project that determines whether the current FAANG metadata standards for livestock can be used to ingest scRNA-seq datasets into Terra in a manner consistent with HCA Data Portal standards. Importantly, rich scRNA-seq metadata can now be brokered through the FAANG data portal using a semi-automated process, thereby avoiding the need for substantial expert curation. We have further extended the functionality of this tool so that validated and ingested SC files within the HCA Data Portal are transferred to Terra for further analysis. In addition, we verified data ingestion into Terra, hosted on Azure, and demonstrated the use of a workflow to analyze the first ingested porcine scRNA-seq dataset. Additionally, we have also developed prototype tools to visualize the output of scRNA-seq analyses on genome browsers to compare gene expression patterns across tissues and cell populations. This JBrowse tool now features distinct tracks, showcasing PBMC scRNA-seq alongside two bulk RNA-seq experiments. Discussion We intend to further build upon these existing tools to construct a scientist-friendly data resource and analytical ecosystem based on Findable, Accessible, Interoperable, and Reusable (FAIR) SC principles to facilitate SC-level genomic analysis through data ingestion, storage, retrieval, re-use, visualization, and comparative annotation across agricultural species.

Genetics & Heredity

TGCM: (T)rait, (G)ene, and (C)rop Growth (M)odel Directed Targeted Gene Characterization in Sorghum (Final Technical Report)

Understanding which genes control important crop traits could help scientists develop better bioenergy and food crops more efficiently. However, plant genomes contain tens of thousands of genes, and testing each one individually is expensive and time-consuming. This project developed computational tools to predict which genes are most likely to matter, allowing researchers to focus their efforts where they will have the greatest impact. This project developed and validated integrated approaches combining machine learning, quantitative genetics, and crop growth modeling to improve the efficiency of functional gene characterization in sorghum (Sorghum bicolor), a critical bioenergy and food security crop. The research addressed a fundamental challenge in plant biology: the majority of genes in plant genomes lack experimentally validated functions, making it difficult to prioritize which genes to study using resource-intensive reverse genetics approaches.

60 APPLIED LIFE SCIENCES

Secure biosystems design in Saccharomyces cerevisiae establishes effective biocontainment strategies and mechanisms of escape

The widespread application of recombinant DNA and synthetic biology approaches for microbial metabolic engineering pursuits has motivated the development of biocontainment strategies, targeting safe and secure deployment of genetically modified microorganisms (GMMs). However, the design rules and mechanistic drivers governing biocontainment efficacy, as well as impacts of biocontainment upon microbial fitness, remain to be comprehensively evaluated, hindering predictive design and application of these strategies. We have developed a platform for high-resolution analysis of a transactivated kill switch in laboratory and industrial strains of Saccharomyces cerevisiae to assess modes of biocontainment escape and establish design rules for development of kill switch systems in diverse microbes. A camphor-regulated, RelE toxin system was systematically deployed to assess the impacts of differential kill switch copy number and ploidy in laboratory vs industrial strains. CRISPR-mediated integration of the biocontainment system at various loci revealed rapid escape events driven, in part, by mutations to both the Cam-transactivator (cam-TA) and RelE toxin. Genetic engineering enabled recapitulation of escape phenotypes, confirming mechanisms of escape and establishing structure-function relationships in the cam-TA system. Interestingly, genomic resequencing of escape mutants also revealed a series of off-target mutations, implicating additional modes of kill switch escape. Multi-copy integration of the kill switch system mitigated these effects by orders of magnitude, without compromising the biosynthetic capacity of the microbes, but proved insufficient to establish sustained biocontainment. The resultant data define a series of key design rules for next-generation biocontainment strategies and add to a growing foundational knowledge base targeting establishment of secure biosystems designs.

59 BASIC BIOLOGICAL SCIENCES

Resilient Crop Production for the Bioeconomy: Frontier Science for the Bioeconomy Workshop Series

A bioeconomy based on energy, chemicals, and bioproducts derived from nonfood crops promotes U.S. prosperity, energy independence, and national security. Crop productivity (i.e., yield) affects profitability for farmers, whereas biomass and seed oil quality affects profitability for biorefineries through the conversion efficiency of feedstock crops into energy, chemicals, and bioproducts. Scientific breakthroughs in plant genetics and genomics, ecological processes, bioprocessing, microbial engineering and design, and conversion technologies have improved the potential market viability of products derived from nonfood crops specifically cultivated as biomass and seed oil feedstocks. Improving understanding of a plant’s complex responses to withstand or recover from stress is crucial to maintaining feedstock yield and quality, particularly in challenging production zones. Stress may be related to biotic or abiotic (i.e., nonliving) environmental conditions that negatively affect plant growth, development, and productivity. Research and development to enhance the environmental resilience of feedstock crops will enable improved production by minimizing losses during stress and improving production on currently economically nonviable marginal lands.

09 BIOMASS FUELS

Integrase-on-Demand

SAND2025-07449O Integrase-on-Demand is a software tool that allows users to identify regions in genomic sequences where genetic material can be integrated with high probability. It uses a database of integrases and their DNA attachment sites to search against any genomic sequence, producing a list of open sites, the integrase sequence, and the source of the genomic island. The program requires MASH software to be available on the system. It consists of a main script and a precomputed input file, with a taxonomy mode that searches closely related genomes and a search mode that looks for identical attachment site matches in the integrase/attachment input file. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Williams, Kelly [Sandia National Lab. (SNL-CA), Li

Status on Genetic Resistance to Rice Blast Disease in the Post-Genomic Era

Rice blast, caused by Magnaporthe oryzae, is a major threat to global rice production, necessitating the development of resistant cultivars through genetic improvement. Breakthroughs in rice genomics, including the complete genome sequencing of japonica and indica subspecies and the availability of various sequence-based molecular markers, have greatly advanced the genetic analysis of blast resistance. To date, approximately 122 blast-resistance genes have been identified, with 39 of these genes cloned and molecularly characterized. The application of these findings in marker-assisted selection (MAS) has significantly improved rice breeding, allowing for the efficient integration of multiple resistance genes into elite cultivars, enhancing both the durability and spectrum of resistance. Pangenomic studies, along with AI-driven tools like AlphaFold2, RoseTTAFold, and AlphaFold3, have further accelerated the identification and functional characterization of resistance genes, expediting the breeding process. Future rice blast disease management will depend on leveraging these advanced genomic and computational technologies. Emphasis should be placed on enhancing computational tools for the large-scale screening of resistance genes and utilizing gene editing technologies such as CRISPR-Cas9 for functional validation and targeted resistance enhancement and deployment. These approaches will be crucial for advancing rice blast resistance, ensuring food security, and promoting agricultural sustainability.

Pedrozo, Rodrigo

A DNA-Inspired Encryption Methodology for Secure, Mobile Ad Hoc Networks

Users are pushing for greater physical mobility with their network and Internet access. Mobile ad hoc networks (MANET) can provide an efficient mobile network architecture, but security is a key concern. A figure summarizes differences in the state of network security for MANET and fixed networks. MANETs require the ability to distinguish trusted peers, and tolerate the ingress/egress of nodes on an unscheduled basis. Because the networks by their very nature are mobile and self-organizing, use of a Public Key Infra structure (PKI), X.509 certificates, RSA, and nonce ex changes becomes problematic if the ideal of MANET is to be achieved. Molecular biology models such as DNA evolution can provide a basis for a proprietary security architecture that achieves high degrees of diffusion and confusion, and resistance to cryptanalysis. A proprietary encryption mechanism was developed that uses the principles of DNA replication and steganography (hidden word cryptography) for confidentiality and authentication. The foundation of the approach includes organization of coded words and messages using base pairs organized into genes, an expandable genome consisting of DNA-based chromosome keys, and a DNA-based message encoding, replication, and evolution and fitness. In evolutionary computing, a fitness algorithm determines whether candidate solutions, in this case encrypted messages, are sufficiently encrypted to be transmitted. The technology provides a mechanism for confidential electronic traffic over a MANET without a PKI for authenticating users.

Shaw, Harry

Atomic Simulation of Complex DNA DSBs and the Interactions with the Ku70/80 Heterodimer

DNA double strand breaks (DSBs) induced by ionizing radiation (IR) usually contain modified bases such as 8-oxo-7,8-dihydroguanine (8-oxoG) and thymine glycol, apurinic/apyrimidinic (AP) sites, 2-deoxyribonolactone, or single-strand breaks (SSBs). The presence of such lesions in close proximity to the DSB terminus makes the DNA nicks more difficult to repair and rejoin than endogenously induced simple DSBs, and as such a major determinant of the biological effects of high linear energy transfer (LET) radiation as encountered in space travel. In this study we conducted molecular dynamics simulations on a series of DNA duplexes with various complex lesions of 8-oxoG and AP sites, in an effort to investigate the effects of such lesions to the structural integrity and stability of DNA after insulted by IR. We also simulated the interaction of such complex DSBs with the Ku70/80 heterodimer, the first protein in mammalian cells to embark the non-homologous end joining (NHEJ) DNA repair pathway. The results indicate, compared to DNA with simple DSBs, the complex lesions can enhance the hydrogen bonds opening rate at the DNA terminus, and increase the mobility of the whole duplex, thus they present more deleterious effects to the genome integrity if not captured and repaired promptly in cells. Simulations also demonstrate the binding of Ku drastically reduces structural disruption and flexibility caused by the complex lesions, and the interactions of Ku with complex DSBs have a different potential energy landscape from the bound structure with simple DSB. In all complex DSBs systems, the binding of DSB terminus with Ku70 is softened while the binding of the middle duplex with Ku80 is tightened. This energy shift may help the Ku protein to secure at the DSB terminus for a longer time, so that other end processing factors or repair pathways can proceed at the lesions before NHEJ repair process starts. These atomic simulations may provide valuable new insight into the selective action of repair proteins on damaged DNA.

Hu, Shaowen

Digital Droplet PCR and Mesocosm-Based Methods to Evaluate Biocontainment Strategies in a Native Soil Ecosystem

Genetically modified industrial production microbes and their associated bioproducts have emerged as an integral component of a sustainable bioeconomy. However, the rapid development of these innovative technologies raises biosecurity concerns, namely, the risk of environmental escape. Thus, the realization of a bioeconomy hinges not only on the development and deployment of microbial production hosts, but also on the development of secure biosystems and biocontainment designs. Current laboratory-based biocontainment testing systems do not accurately reflect the complexities found in natural environments, necessitating an environmentally relevant analysis pipeline that allows for the detection of rare escapees within a complex soil microbiome and differentiation between closely related strains. To this end, we have developed an approach that utilizes soil mesocosms and integrated digital droplet PCR (ddPCR) system to evaluate the efficacy of novel biocontainment strategies. We demonstrate the utility of this approach by modeling contamination with industrial microbial chasses versus their biocontained counterparts. Here we demonstrate the broad utility of this system by highlighting findings from strains of Saccharomyces cerevisiae that are contained with an inducible toxin anti-toxin system, strains of Synechocystis sp. PCC 6803 contained via gene knockout or toxin anti-toxin system, and strains of Escherichia coli that are contained via genomic recoding. We also show that ddPCR can be used to detect gene copies from E. coli equal to those counted by traditional spot plating assays. The resultant data demonstrates that this system has broad utility across diverse microbial chassis and biocontainment strategies and enables researchers to track the fate of our contaminating microbe with high sensitivity in the soil. The findings presented here support the use of this mesocosm-based approach to assess the environmental impact of industrial microbes and to validate biocontainment strategies.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Poplar

SAND2025-00683O Poplar is a software tool that generates a phylogenetic tree from input gene and genome sequences. It integrates established tools to identify genes within genomes, group sequences, construct gene trees, and infer a species tree. Poplar processes nucleotide sequences, identifies similar sequences using Nucleotide BLAST, groups them with DBSCAN, aligns sequences with MAFFT, constructs gene trees with RAxML-NG, and infers a species tree using ASTRAL-Pro3. This pipeline provides a structured approach to phylogenetic analysis, facilitating the study of evolutionary relationships among species. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Krishnakumar, Raga

Genomic Surveillance Detection of SARS-CoV-1–Like Viruses in Rhinolophidae Bats, Bandarban Region, Bangladesh

We sequenced sarbecovirus from Rhinolophus spp. bats in Bandarban District, Bangladesh, in a genomic surveillance campaign during 2022–2023. Sequences shared identity with SARS-CoV-1 Tor2, which caused an outbreak of human illnesses in 2003. Describing the genetic diversity and zoonotic potential of reservoir pathogens can aid in identifying sources of future spillovers.

Angiotensin Converting Enzyme 2

Selection Factors for Space Crops

NASA is actively researching space crop production to determine its potential to contribute to food system security on long duration missions beyond Low Earth Orbit. Our near-term focus is on nutrient and variety supplementation of prepackaged food with fresh produce that requires little or no processing. The longer-term goal is caloric replacement to become less dependent on Earth, and this will require cultivation of staple crops, processing and cooking equipment, integration with spacecraft air, water, and power systems, and automation. There are numerous technology and knowledge gaps remaining for sustainable space crop production systems, but one high-impact area is in the development of crops specifically customized to meet the needs of controlled environment crop production, astronaut health and well-being, and space-unique environments. Modern crop breeding and genome engineering tools are allowing for rapid development of new genotypes with incredible specificity. Targeted aspects to optimize crops for space have been identified and characterized into five categories: plant growth and development, plant physiology, produce nutrition, produce organoleptic acceptability, and postharvest characteristics. Within each category there are several targets that further the development of crop production systems for spaceflight, such as crop size and harvest index, tolerance to specific environmental stresses, optimizing target nutrients that are low or degrade in the packaged diet, maintenance time requirements, and less indigestible structural material. NASA-funded PIs are already beginning to develop candidate crops, and spaceflight testing and validation of novel space crops is on the horizon. Crops developed for space also have the potential to benefit terrestrial controlled environment agriculture crop production systems. This research was supported by NASA’s Space Biology and Human Research Programs.

Space Crop Production

Selection Factors for Space Crops

NASA is actively researching space crop production to determine its potential to contribute to food system security on long duration missions beyond Low Earth Orbit. Our near-term focus is on nutrient and variety supplementation of prepackaged food with fresh produce that requires little or no processing. The longer-term goal is caloric replacement to become less dependent on Earth, and this will require cultivation of staple crops, processing and cooking equipment, integration with spacecraft air, water, and power systems, and automation. There are numerous technology and knowledge gaps remaining for sustainable space crop production systems, but one high-impact area is in the development of crops specifically customized to meet the needs of controlled environment crop production, astronaut health and well-being, and space-unique environments. Modern crop breeding and genome engineering tools are allowing for rapid development of new genotypes with incredible specificity. Targeted aspects to optimize crops for space have been identified and characterized into five categories: plant growth and development, plant physiology, produce nutrition, produce organoleptic acceptability, and postharvest characteristics. Within each category there are several targets that further the development of crop production systems for spaceflight, such as crop size and harvest index, tolerance to specific environmental stresses, optimizing target nutrients that are low or degrade in the packaged diet, maintenance time requirements, and less indigestible structural material. NASA-funded PIs are already beginning to develop candidate crops, and spaceflight testing and validation of novel space crops is on the horizon. Crops developed for space also have the potential to benefit terrestrial controlled environment agriculture crop production systems. This research was supported by NASA’s Space Biology and Human Research Programs.

Space Crop Production