Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “data enhancement”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Dynamic CCS-EJ-SJ Database and Web Application - What's New

At the 2024 FECM/NETL Carbon Management Research Project Review Meeting, within the Carbon Transport and Storage Breakout Session 3, the presentation "Dynamic CCS-EJ-SJ Database and Web Application - What's New" highlights the critical tool designed to integrate environmental and social justice considerations into Carbon Capture and Storage (CCS) projects. Key features include an interactive dashboard for data access and visualization, which supports stakeholders in making informed decisions regarding CCS implementation, and updated data layers. The latest version enhances data integration and usability, providing a comprehensive resource for assessing the social and environmental impacts of CCS projects. There are 7 categories in the CCS EJSJ v2 database (released 03/31/2024): environmental justice, energy justice, economic justice, social justice, ecosystem assets, clean energy, and infrastructure. Most of the layers within each category have been updated in this version. As compared to the old database, there are 3 new categories in the v2 database: ecosystem assets, clean energy, and infrastructure.

Sharma, Maneesh↗

University Data Management Pilot Utilizing the Nuclear Research Data System

Background In 2022, the Office of Science and Technology Policy (OSTP) issued a memo that significantly reshaped the landscape of access to federally funded research. The memo mandated that all taxpayer-funded research be made available to the public without delay upon publication, without an embargo period, superseding the 2013 OSTP public access policy. This public access policy promotes transparency and the democratization of knowledge, ensuring that the fruits of scientific endeavors funded by federal agencies could be immediately accessed and built upon by scientists, educators, students, and the public at large. To implement the requirements of the OSTP guidance and DOE Public Access Plan, the Office of Nuclear Energy (NE) has implemented public access plan guidance and has identified several areas where better data management practices would further expand public access to important nuclear energy related scientific data, reports, and other technical products. Significant NE supported efforts are already underway for data management and public access to important nuclear energy related data.1 2 To address gaps in data management practices, and improve retention and accessibility of data, NE is actively exploring enhanced data management options utilizing its high-performance computing resources administered by its Nuclear Scientific User Facility Program. A newly piloted system, the Nuclear Research Data System (NRDS) acts as a portal for data collection and dissemination. Nuclear Energy University Program Research and Development Portfolio According to Web of Science, NEUP has produced 2,345 journal publication that have been cited more than 61,000 times3 and countless conference proceedings. These publications are publicly available through OSTI.gov and in the open literature. Additional scientific and technical products including project milestones that are not publications and NEUP project final reports are vetted through OSTI.gov and released once reviewed and approved by DOE. Since 2009, NEUP has awarded close to 1,000 different R&D projects in technical areas across the NE research programs. As of June 2023, 512 NEUP reports are publicly available on OSTI. The underlying data for projects is still held at universities, and data transfer, co-location, and dissemination has not occurred in a systematic way. NEUP data is currently accessible through myriad university-based data repositories, or through direct requests to PIs. The program identified this patchwork of repositories, or often lack of publicly available data, as a significant barrier to an organized, accessible, and comprehensive solution to sharing data with the larger nuclear energy community. Approach The goal of this pilot project is to establish a pathway to a consolidated long-term repository for NEUP project data. To accomplish this goal, the pilot strives to accomplish the following objectives: Establish data collection standards, including a standard set of required supplementary information to contextualize and support raw data files. Work with the HPC group collect and upload information and to modify the NRDS system, as needed, to support a standardized approach. Resolve potential barriers to successful roll out of an expanded data collection strategy, including modifying data management plan guidelines and establishing a document and data release process that accounts for potential intellectual property and/or export control concerns. Results Overall, the pilot was successful in collecting 8,982 raw and processes data files, 220 reports, 56 calibration files, and 5,931 other supplementary documents. Supplementary documents included experimental plans, methods, journal publications and conference proceedings, milestone reports, and final reports. Figure 2 shows the number of data sets and supplementary project information provided by each project. Projects has significantly different input, depending on experimental data produced and completeness of the datasets provided.

Data collection↗

Advanced Interactive 3D Visualization Tool for Customizable Analyses of Tomography Datasets in Material Science

Current methods for visualizing and analyzing 3D tomography datasets in materials science often lack the interactivity and depth required for detailed structural insights. This limitation restricts a researchers' ability to accurately interpret complex data, which is critical for advancing material innovations and understanding structural properties. To address this issue, we have developed a novel, web-based interactive 3D visualization and analysis tool from the Trame framework that offers customizable features to enhance data interpretability. The tool allows users to adjust parameters such as visible range, slice planes, data rotation, and layering, providing a more detailed and dynamic view of complex structures. Its user-friendly web interface increases the accessibility and ease of use for both novice and experienced researchers, to visualize large volumetric datasets. The tool supports a diverse range of data formats, making it versatile for various research applications. Unique capabilities include real-time data manipulation, automated feature detection, context-sensitive feedback, and real-time volume calculations and distributions per sliced region or layer, alongside the ability to quickly generate high-quality screenshots and videos for presentations and reports. These advancements offer a comprehensive solution for enhanced 3D data exploration, significantly improving the analysis process and communication of results in materials science.

36 - MATERIALS SCIENCE↗

The Monarch Initiative in 2024: an analytic platform integrating phenotypes, genes and diseases across species

Abstract Bridging the gap between genetic variations, environmental determinants, and phenotypic outcomes is critical for supporting clinical diagnosis and understanding mechanisms of diseases. It requires integrating open data at a global scale. The Monarch Initiative advances these goals by developing open ontologies, semantic data models, and knowledge graphs for translational research. The Monarch App is an integrated platform combining data about genes, phenotypes, and diseases across species. Monarch's APIs enable access to carefully curated datasets and advanced analysis tools that support the understanding and diagnosis of disease for diverse applications such as variant prioritization, deep phenotyping, and patient profile-matching. We have migrated our system into a scalable, cloud-based infrastructure; simplified Monarch's data ingestion and knowledge graph integration systems; enhanced data mapping and integration standards; and developed a new user interface with novel search and graph navigation features. Furthermore, we advanced Monarch's analytic tools by developing a customized plugin for OpenAI’s ChatGPT to increase the reliability of its responses about phenotypic data, allowing us to interrogate the knowledge in the Monarch graph using state-of-the-art Large Language Models. The resources of the Monarch Initiative can be found at monarchinitiative.org and its corresponding code repository at github.com/monarch-initiative/monarch-app.

60 APPLIED LIFE SCIENCES↗

Energy efficient data center liquid cooling with geothermal enhancement

A data center cooling system is operated in a first mode, has an indoor portion wherein heat is absorbed from components in the data center by a heat transfer fluid, and has an outdoor heat exchanger portion and a geothermal heat exchanger portion. The first mode includes ambient air cooling of the heat transfer fluid in the outdoor heat exchanger portion and/or geothermal cooling of the heat transfer fluid in the geothermal heat exchanger portion. Based on an appropriate metric, a determination is made that a switch should be made from the first mode to a second, different, mode; and, responsive thereto, the data center cooling system is switched to the second mode. The second mode includes at least another of ambient air cooling of the heat transfer fluid in the outdoor heat exchanger portion and geothermal cooling of the heat transfer fluid in the geothermal heat exchanger portion.

Chainer, Timothy J.↗

Quality Guidelines for Energy System Studies: Process Modeling Design Parameters

The National Energy Technology Laboratory (NETL) conducts systems analysis studies that require a large number of inputs, from ambient conditions to parameters for Aspen Plus ® (Aspen) process blocks. The sheer number of assumptions required makes it impractical to document all of them in each issued report. The purpose of the Quality Guidelines for Energy System Studies (QGESS) is to document the assumptions most commonly used in system analysis studies and the basis for those assumptions. In order to develop the systems analysis models presented in various NETL reports, significant vendor data have been obtained, and these data enhance the model outputs. Much of the vendor data obtained are considered proprietary and not suitable for public release or attribution to a specific vendor. As such, several sub-systems common in NETL reports and their process parameter data are not reported in this document to protect proprietary vendor information. The values and ranges of values presented in this report represent assumptions that have been made in previous studies.

97 MATHEMATICS AND COMPUTING↗

Community Requirements Meta-Analysis: Characterizing Needs and Opportunities for HPDF

This High Performance Data Facility (HPDF) Project is creating a new scientific user facility to provide advanced infrastructure for data-intensive science, supporting the DOE’s Office of Science (SC) community. HPDF’s mission is to enable and accelerate scientific discovery by delivering state-of-the-art data management infrastructure, capabilities, and tools. This meta-analysis examines the needs of the breadth of the SC community, captured in publicly available community reports or mission documents. The meta-analysis identifies and provides initial characterization of fifteen core requirements for the HPDF Project team to consider during the conceptual design phase. The fifteen requirements illustrate how scientific work among SC communities requires modern, seamless user experiences across the ASCR Ecosystem to advance the use of large volumes of heterogeneous data. The scientific community requires support for the missing middle of compute between local and HPC to interactively and collaboratively use growing datasets. Data producers and end users will benefit from enhanced data catalogs and portals that improve data access through advanced search of well curated data. The fifteen requirements are examined here organized across five themes for discussion. Examples in each theme illustrate the array of scientific needs that convey the important role that the fully realized and operational High Performance Data Facility will be able to play as an integral part of the evolving ASCR Ecosystem. Our amalgamated data tables from ESnet reports demonstrate ranges to the volumes of data HPDF must be concerned with, but limitations are inherent to this meta-analysis (see Key Challenges & Limitations). Feedback and validation of these requirements along with additional details and emergent community requirements will be gathered through user research and design activities.

97 MATHEMATICS AND COMPUTING↗

TRANSP integrated modeling code for interpretive and predictive analysis of tokamak plasmas

This paper provides a comprehensive review of the TRANSP code, a sophisticated tool for interpretive and predictive analysis of tokamak plasmas, detailing its major capabilities and features. It describes the equations for particle, power, and momentum balance analysis, as well as the poloidal field diffusion equations. The paper outlines the spatial and time grids used in TRANSP and details the equilibrium assumptions and solvers. Various models for heating and current drive and radiation, including updates to the NUBEAM model, are discussed. The handling of large-scale events such as sawtooth crashes and pellet injections is examined, along with the predictive capabilities for advancing plasma profiles. The integration of TRANSP with the ITER Integrated Modeling and Analysis Suite (IMAS) is highlighted, demonstrating enhanced data access and analysis capabilities. Additionally, the paper discusses best practices and continuous integration techniques to enhance TRANSP's robustness. The suite of TRANSP tools, designed for efficient data analysis and simulation, further supports the optimization of tokamak operations and coupling with other tokamak codes. Continuous development and support ensure that TRANSP remains a major code for the analysis of experimental data for controlled thermonuclear fusion, maintaining its critical role in supporting the optimization of tokamak operations and advancing fusion research.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

HPDR: High-Performance Portable Scientific Data Reduction Framework

The rapid growth in scientific data generation is outpacing advancements in computing systems necessary for efficient storage, transfer, and analysis, particularly in the context of exascale computing. With the deployment of first-generation exascale computing systems and next-generation experimental facilities, this gap is widening and necessitates effective data reduction techniques to manage enormous data volumes. Over the past decade, various data reduction methods, including lossless compression, error-controlled lossy compression, and data refactoring, have been developed to accelerate I/O in scientific workflows. Despite significant reductions in data volume, these methods introduce considerable computational overhead, which can become the new bottleneck in data processing. To mitigate this, GPU-accelerated data reduction algorithms have been introduced. However, challenges remain in their integration into exascale workflows, including limited portability across different GPU architectures, substantial memory transfer overhead, and reduced scalability on dense multi-GPU systems. To address these challenges, we propose HPDR, a high-performance and portable data reduction framework. HPDR is designed to enable the execution of state-of-the-art reduction algorithms across diverse processor architectures while reducing memory transfer overhead to 2.3 % of the original, resulting in up to 3.5× faster throughput compared to existing solutions. It also achieves up to 96% of the theoretical speedup in multi-GPU settings. In addition, evaluations on accelerating I/O operations at scale up to 1,024 nodes of the Frontier supercomputer demonstrate that HPDR can achieve up to 103 TB/s reduction throughput, providing up to 4× acceleration in parallel I/O performance compared to existing data reduction routines. This work highlights the potential of HPDR to significantly enhance data reduction efficiency in exascale computing environments.

Chen, Jieyang [University of Oregon]↗

AI-Driven Crack Detection for Remanufacturing Cylinder Heads Using Deep Learning and Engineering-Informed Data Augmentation

Detecting cracks in cylinder heads traditionally relies on manual inspection, which is time-consuming and susceptible to human error. As an alternative, automated object detection utilizing computer vision and machine learning models has been explored. However, these methods often face challenges due to a lack of sufficiently annotated training data, limited image diversity, and the inherently small size of cracks. Addressing these constraints, this paper introduces a novel automated crack-detection method that enhances data availability through a synthetic data generation technique. Unlike general data augmentation practices, our method involves copying cracks from one location to another, guided by both random and informed engineering decisions about likely crack formations due to cyclic thermomechanical loads. The innovative aspect of our approach lies in the integration of domain-specific engineering knowledge into the synthetic generation process, which substantially improves detection accuracy. We evaluate our method’s effectiveness using two metrics: the F2 score, which emphasizes recall to prioritize detecting all potential cracks, and mean average precision (MAP), a standard measure in object detection. Experimental results demonstrate that, without engineering insights, our method increases the F2 score from 0.40 to 0.65, while maintaining a stable MAP. Incorporating detailed engineering knowledge further enhances the F2 score to 0.70 and improves MAP to 0.57, representing increases of 63% and 43%, respectively. These results confirm that our approach not only mitigates the limitations of traditional data augmentation but also significantly advances the reliability and precision of crack detection in industrial settings.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Scale-up Unlearnable Examples Learning with High-performance Computing

Recent advancements in AI models, like ChatGPT, are structured to retain user interactions, which could inadvertently include sensitive healthcare data. In the healthcare field, particularly when radiologists use AI-driven diagnostic tools hosted on online platforms, there is a risk that medical imaging data may be repurposed for future AI training without explicit consent, spotlighting critical privacy and intellectual property concerns around healthcare data usage. Addressing these privacy challenges, a novel approach known as Unlearnable Examples (UEs) has been introduced, aiming to make data unlearnable to deep learning models. A prominent method within this area, called Unlearnable Clustering (UC), has shown improved UE performance with larger batch sizes but was previously limited by computational resources (e.g., a single workstation). To push the boundaries of UE performance with theoretically unlimited resources, we scaled up UC learning across various datasets using Distributed Data Parallel (DDP) training on the Summit supercomputer. Our goal was to examine UE efficacy at high-performance computing (HPC) levels to prevent unauthorized learning and enhance data security, particularly exploring the impact of batch size on UE’s unlearnability. Utilizing the robust computational capabilities of the Summit, extensive experiments were conducted on diverse datasets such as Pets, MedMNist, Flowers, and Flowers102. Our findings reveal that both overly large and overly small batch sizes can lead to performance instability and affect accuracy. However, the relationship between batch size and unlearnability varied across datasets, highlighting the necessity for tailored batch size strategies to achieve optimal data protection. The use of Summit’s high-performance GPUs, along with the efficiency of the DDP framework, facilitated rapid updates of model parameters and consistent training across nodes. Our results underscore the critical role of selecting appropriate batch sizes based on the specific characteristics of each dataset to prevent learning and ensure data security in deep learning applications. The source code is publicly available at https: // github. com/ hrlblab/ UE_ HPC .

Zhu, Yanfan [Vanderbilt University, Nashville, TN,↗

AI-driven neutrino diagnostics and radiation-hard beam instrumentation for next-generation neutrino experiments

The Long Baseline Neutrino Facility (LBNF) at Fermilab will deliver a high-intensity, multi-megawatt neutrino beam to the Deep Underground Neutrino Experiment (DUNE), enabling precision tests of the three-neutrino paradigm, CP violation searches, neutrino mass ordering determination, and supernova neutrino studies. To accelerate DUNE’s physics reach and ensure robust beam operations, we propose an integrated AI-driven framework with real-time diagnostics and radiation-hardened instrumentation. At its core is a Real-Time Beam Integrity Monitor using a physics-informed Digital Twin. By reconstructing pion phase space from muon profiles and exploiting magnetic horn optic linearity, it enables spill-by-spill beam correction and flux stabilization. By using this approach, flux-related systematics could be reduced from 5% to 1%, potentially accelerating the discovery of CP violations by four to six years. Complementing this, a US–Japan R&D effort will deploy a LAPPD-based muon monitor in the NuMI beamline. ToF measurements can be acquired with picosecond precision using this radiation-hard system, enhancing sensitivity to horn chromatic effects. Simulations confirm strong response to these effects. ML models predict beam quality and horn current to sub-percent accuracy from muon data, enhancing anomaly detection and stability. This scalable, AI-enabled strategy improves beam fidelity, reduces systematics, and sets a new standard for high-power accelerator operations.

Ganguly, Sudeshna [Fermilab] (ORCID:00000003163482↗

AI-driven neutrino diagnostics and radiation-hard beam instrumentation for next-generation neutrino experiments

The Long Baseline Neutrino Facility (LBNF) at Fermilab will deliver a high-intensity, multi-megawatt neutrino beam to the Deep Underground Neutrino Experiment (DUNE), enabling precision tests of the three-neutrino paradigm, CP violation searches, neutrino mass ordering determination, and supernova neutrino studies. To accelerate DUNE’s physics reach and ensure robust beam operations, we propose an integrated AI-driven framework with real-time diagnostics and radiation-hardened instrumentation. At its core is a Real-Time Beam Integrity Monitor using a physics-informed Digital Twin. By reconstructing pion phase space from muon profiles and exploiting magnetic horn optic linearity, it enables spill-by-spill beam correction and flux stabilization. By using this approach, flux-related systematics could be reduced from 5% to 1%, potentially accelerating the discovery of CP violations by four to six years. Complementing this, a US–Japan R&D effort will deploy a LAPPD-based muon monitor in the NuMI beamline. ToF measurements can be acquired with picosecond precision using this radiation-hard system, enhancing sensitivity to horn chromatic effects. Simulations confirm strong response to these effects. ML models predict beam quality and horn current to sub-percent accuracy from muon data, enhancing anomaly detection and stability. This scalable, AI-enabled strategy improves beam fidelity, reduces systematics, and sets a new standard for high-power accelerator operations.

Ganguly, Sudeshna [Fermilab] (ORCID:00000003163482↗

The state of the art for neutron irradiation experiments from the perspective of the High Flux Isotope Reactor (HFIR)

Irradiation experiment campaigns are critical to advancing nuclear energy technologies by providing data on material performance under relevant radiation conditions. Successful irradiation experiments require integrated design efforts that balance technical goals with facility constraints. Here, this paper presents an expert-informed overview of irradiation experiment design at the High Flux Isotope Reactor. It addresses the nuclear materials research and irradiation experiment communities to guide them toward developing technically sound, facility-compatible campaigns. The High Flux Isotope Reactor is a multipurpose reactor supporting isotope production, neutron scattering, and materials testing. Its high, steady-state neutron flux is ideal for irradiation experiments, but successful execution demands coordinated thermal, structural, and reactor physics analyses. The paper outlines the complete development workflow from concept definition and design optimization to safety qualification and post-irradiation examination. Standardized capsule platforms are also discussed in terms of flexibility, specimen capacity, and thermal performance. Common failure modes such as unanticipated geometric variations, can impact temperature-dose profiles and compromise data reliability. Therefore, detailed thermal modeling and accurate as-built characterization are essential for meaningful post-irradiation data interpretation. Key recommendations include early engagement all stakeholders, clearly defined design expectations, and alignment of specimen geometries with post-irradiation examination capabilities. This approach reduces design iterations, enhances data quality, and supports more efficient use of irradiation resources. Strategic and well-planned irradiation testing not only improves individual campaign success but also accelerates the deployment of advanced nuclear technologies. By closing critical data gaps and reducing development risks, the nuclear materials community can more effectively contribute to the future of clean, resilient energy systems.

Experiments↗

Electrical Fault and Power Quality Detection Algorithms and Customer-Owned DERs Monitoring with a Cyber Grid Guard System and DLT

In this study, the electrical fault and power quality detection algorithms and customer-owned DERs monitoring use cases were implemented, with a Cyber Grid Guard system and DLT. Electrical utilities continue to deploy more intelligent electronic devices (IEDs) inside and outside electrical substations, and are associated with customer-owned distributed energy resources (DERs). Data from these IEDs, such as power meters and protection relays, must be kept confidential and of high integrity. Blockchain technology has the potential to increase microgrid resilience by enhancing data sharing security. The growing use of IEDs and customer-owned renewable energy sources (DERs) may make it necessary to connect Distributed Ledger Technology (DLT) with power system applications. We implemented the electrical faulted phase detection and power quality monitoring algorithms with a Cyber Grid Guard (CGG) system using DLT. In addition, the DERs (wind turbine farms) use case and protective relay cyber-event tests were assessed, by using the CGG system with DLT. In the experimental model, the testbed was created by using a real-time simulator and CGG system with power meters/ protective relays in the loop. The data collected from the CGG system and IEDs were compared with the same time stamp source. These results showed the successful assessment of protection, control and monitoring applications using a CGG system with DLT. In the future, power system applications for the ESGT with DERs and the CGG system will be based on executing smart contracts between electrical utilities and customer-owned DERs.

Piesciorovsky, Emilio↗

Time for a drought experiment: Do you know your plants’ water status?

Abstract Drought stress is an increasing concern because of climate change and increasing demands on water for agriculture. There are still many unknowns about how plants sense and respond to water limitation, including which genes and cellular mechanisms are impactful for ecology and crop improvement in drought-prone environments. A better understanding of plant drought resistance will require integration of several research disciplines. A common set of parameters to describe plant water status and quantify drought severity can enhance data interpretation and research integration across the research disciplines involved in understanding drought resistance and would be especially useful in integrating the flood of genomic data being generated in drought studies. Water potential (ψw) is a physical measure of the free energy status of water that, along with related physiological measurements, allows unambiguous description of plant water status that can apply across various soil types and environmental conditions. ψw and related physiological parameters can be measured with relatively modest investment in equipment and effort. Thus, we propose that increased use of ψw as a fundamental descriptor of plant water status can enhance the insight gained from many drought-related experiments and facilitate data integration and sharing across laboratories and research disciplines.

Juenger, Thomas E. (ORCID:0000000195509288)↗