Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “data enhancement”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

The Upgrade of Horizon-T Detector

The Horizon-T experiment is located at the elevation of 3346 m above sea level near the city of Almaty, Republic of Kazakhstan. A thorough comparison of the spatial and temporal characteristics of charged components of Extended Air Showers (EAS) with delayed particles has been conducted between the simulated EAS using CORSIKA simulation package and the selection from the experimental data set of events with two pulses recorded by a detector at ~600 m distance from axis [1]. This comparison has shown that events with delayed particles cannot be described within existing simulation models. The significance of these results prompted the upgrade of the Horizon-T experiment. New detector points have been added at the ~600m to enhance data at that distance. Fast glass-based detector has been added to the scintillator-based detector at center point for accurate measurements of the pulse widths.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Aligning Standards Communities for Omics Biodiversity Data: Sustainable Darwin Core-MIxS Interoperability

The standardization of data, encompassing both primary and contextual information (metadata), plays a pivotal role in facilitating data (re-)use, integration, and knowledge generation. However, the biodiversity and omics communities, converging on omics biodiversity data, have historically developed and adopted their own distinct standards, hindering effective (meta)data integration and collaboration. In response to this challenge, the Task Group (TG) for Sustainable DwC-MIxS Interoperability was established. Convening experts from the Biodiversity Information Standards (TDWG) and the Genomic Standards Consortium (GSC) alongside external stakeholders, the TG aimed to promote sustainable interoperability between the Minimum Information about any (x) Sequence (MIxS) and Darwin Core (DwC) specifications. To achieve this goal, the TG utilized the Simple Standard for Sharing Ontology Mappings (SSSOM) to create a comprehensive mapping of DwC keys to MIxS keys. This mapping, combined with the development of the MIxS-DwC extension, enables the incorporation of MIxS core terms into DwC-compliant metadata records, facilitating seamless data exchange between MIxS and DwC user communities. Through the implementation of this translation layer, data produced in either MIxS- or DwC-compliant formats can now be efficiently brokered, breaking down silos and fostering closer collaboration between the biodiversity and omics communities. To ensure its sustainability and lasting impact, TDWG and GSC have both signed a Memorandum of Understanding (MoU) on creating a continuous model to synchronize their standards. These achievements mark a significant step forward in enhancing data sharing and utilization across domains, thereby unlocking new opportunities for scientific discovery and advancement.

59 BASIC BIOLOGICAL SCIENCES↗

CRiSPPy: An advanced hydropower scheduling tool for the Colorado River Storage Project

The Western Area Power Administration (WAPA) plays a vital role in delivering reliable and cost-effective hydroelectric power to millions of customers across the western United States. The Colorado River Storage Project (CRSP) carries out WAPA’s mission in Arizona, Utah, Colorado, New Mexico, Nevada, Wyoming and Texas. Achieving this mission requires effective management of the Colorado River system, and depends on the use of advanced analytical tools and modeling methodologies. For many years, CRSP has relied on the Generation and Transmission Maximization Superlite (GTMax SL) model for its mid-term and long-term hydroscheduling needs. However, the evolving energy market, power system operations, environmental rules, and hydrology conditions, coupled with advancements in computational capabilities, have necessitated the development of a more modern and robust solution. This report introduces the Colorado River Storage Project Python-based (CRiSPPy) model, a new, advanced hydropower scheduling tool developed to address CRSP ever-evolving challenges. CRiSPPy represents a significant leap forward in our ability to model and optimize the operation of the Colorado River system. It incorporates state-of-the-art optimization algorithms, enhanced data management capabilities, and an advanced graphical user interface, providing WAPA CRSP personnel with unprecedented insights and decision-making support. This document details the development, capabilities, and implementation of CRiSPPy. It is intended to serve as a comprehensive resource for WAPA staff, stakeholders, and anyone interested in the future of hydropower scheduling in the Colorado River Basin. We are confident that CRiSPPy will enhance WAPA's mission while adapting to the challenges of a dynamic and increasingly complex environment. The version of CRiSPPy described in this report is the version 2.3. New versions of CRiSPPy will be developed as the tool keeps evolving to address CRSP challenges.

13 HYDRO ENERGY↗

Towards Next-Generation Urban Decision Support Systems through AI-Powered Construction of Scientific Ontology Using Large Language Models—A Case in Optimizing Intermodal Freight Transportation

The incorporation of Artificial Intelligence (AI) models into various optimization systems is on the rise. However, addressing complex urban and environmental management challenges often demands deep expertise in domain science and informatics. This expertise is essential for deriving data and simulation-driven insights that support informed decision-making. In this context, we investigate the potential of leveraging the pre-trained Large Language Models (LLMs) to create knowledge representations for supporting operations research. By adopting ChatGPT-4 API as the reasoning core, we outline an applied workflow that encompasses natural language processing, Methontology-based prompt tuning, and Generative Pre-trained Transformer (GPT), to automate the construction of scenario-based ontologies using existing research articles and technical manuals of urban datasets and simulations. From these ontologies, knowledge graphs can be derived using widely adopted formats and protocols, guiding various tasks towards data-informed decision support. The performance of our methodology is evaluated through a comparative analysis that contrasts our AI-generated ontology with the widely recognized pizza ontology, commonly used in tutorials for popular ontology software. We conclude with a real-world case study on optimizing the complex system of multi-modal freight transportation. Our approach advances urban decision support systems by enhancing data and metadata modeling, improving data integration and simulation coupling, and guiding the development of decision support strategies and essential software components.

96 KNOWLEDGE MANAGEMENT AND PRESERVATION↗

Privacy-Preserving Real-Time Action Detection in Intelligent Vehicles Using Federated Learning-Based Temporal Recurrent Network

This study introduces a privacy-preserving approach for the real-time action detection in intelligent vehicles using a federated learning (FL)-based temporal recurrent network (TRN). This approach enables edge devices to independently train models, enhancing data privacy and scalability by eliminating central data consolidation. Our FL-based TRN effectively captures temporal dependencies, anticipating future actions with high precision. Extensive testing on the Honda HDD and TVSeries datasets demonstrated robust performance in centralized and decentralized settings, with competitive mean average precision (mAP) scores. The experimental results highlighted that our FL-based TRN achieved an mAP of 40.0% in decentralized settings, closely matching the 40.1% in centralized configurations. Notably, the model excelled in detecting complex driving maneuvers, with mAPs of 80.7% for intersection passing and 78.1% for right turns. These outcomes affirm the model’s accuracy in action localization and identification. The system showed significant scalability and adaptability, maintaining robust performance across increased client device counts. The integration of a temporal decoder enabled predictions of future actions up to 2 s ahead, enhancing the responsiveness. Our research advances intelligent vehicle technology, promoting safety and efficiency while maintaining strict privacy standards.

33 ADVANCED PROPULSION SYSTEMS↗

Investigating Temperature Uniformity and Accuracy in PV Module Lamination: A Verification Study

This study investigates the temperature uniformity and accuracy of a photovoltaic (PV) module lamination process by addressing inconsistencies identified in 2017 data where irregular temperature changes were observed across setpoints. The 2017 data showed a notable drop in temperature upon bladder initiation, except for the 145 degrees Celsius profile. This inconsistency indicated potential inaccuracies in manual data recording methods. To address this concern, a verification experiment was conducted to evaluate temperature uniformity across the 2014 Bent River SPL2828 laminator platen and within test samples. Thermocouples, paired with Omega data acquisition software, were deployed to measure temperatures at multiple platen locations and within test samples. The experiment compared lamination temperatures of polyethylene-co-vinyl acetate (EVA) encapsulant when paired with solite glass or TPE backsheets. The methodology included verifying temperature uniformity directly on the platen and by using a large glass/EVA/glass sample using multiple thermocouples. Smaller samples were built with glass/EVA/glass and glass/EVA/backsheet configurations with one centered thermocouple to verify and compare sample temperatures. This verification aims to refine lamination temperature profiles, enhance data accuracy and provide insights into optimal process control for uniform module lamination. Ensuring consistent and uniform lamination may improve the accuracy and reliability of research outcomes.

14 SOLAR ENERGY↗

A Collaboration Website for Muon Catalyzed Fusion and Muon Beam Production

This project sets up a website to support the nascent Muon Catalyzed Fusion collaboration including development of particle accelerators and transport beamlines for muon beams. The website is envisioned as having the general public information pages and private pages for collaboration members. Multimedia elements like images, text animations, and video lectures, covering a broad spectrum of topics will populate the educational site, covering muon facilities, to comprehensive explorations and seminal documents that define the science of Muon Catalyzed Fusion, Acceleration, Applications, Instrumentation, Beamline Design, and beam dynamics design codes. Ensuring compatibility across devices and operating systems, it also features integration with Google Docs for collaboration, a code repository (GitHub), a blog platform with comments (WordPress), the potential for ChatGPT integration and interactive graph plotting with Python Plotty to enhance data visualization. This project will maintain public and protected private pages, due to the proprietary nature of the work or research in progress. The public sections will be built to foster dissemination of information and highlight recent work within the NK Labs collaboration, including lectures, published papers, and regular blog posts with open commenting. The private section will support unpublished or nonpublic research, by facilitating collaborative efforts through integrated Google Docs and Python Plotty for shared graphing work. Ultimately, this project strives to make complex scientific knowledge more accessible to the public, foster enhanced collaboration, and serve as a platform for sharing cutting-edge research in Muon Catalyzed Fusion and Accelerators.

43 PARTICLE ACCELERATORS↗

Simulation of hydropower at subcontinental to global scales: a state-of-the-art review

Abstract Hydroelectric power is playing a new and often expanded role in the world’s major power grids, offering low carbon generating capacity in industrializing, dam-building economies while providing reserve and flexibility to co-manage fledgling wind and solar resources in high income countries. Driven by river flows, conventional hydropower is exposed to the vagaries of weather and climate, motivating drought and climate change hydropower impact studies at large spatial scales. Here we review methods of climate-driven hydropower simulation at large spatial scales, specifically multi-basin regions to global. We identify four types of approach based on complexity of tools and richness of data applied to the problem. Since the earliest attempts to model climate-driven hydropower at continental scale almost two decades ago, the field has transitioned from one of scientific curiosity to practical application, with studies increasingly motivated by the need to inform power grid expansion planning and operation. As the hydrological and water management models used in large-scale hydropower studies become more sophisticated, new opportunities will emerge to study the impacts of changing hydropower on power system reliability and performance at large power grid scale. To grasp these opportunities, the water resources community must continue to enhance data and models for representing river flows and anthropogenic water use and management at subcontinental to global scales.

13 HYDRO ENERGY↗

Cataloging Legacy Data from the Tritium Systems Test Assembly Program

The Tritium Systems Test Assembly (TSTA) at Los Alamos National Laboratory, operational from 1984 to 2001, was critical in advancing fusion fuel cycle technologies, including tritium storage, gas separation, and pumping. TSTA’s contributions, particularly in safe tritium operations, have influenced subsequent fusion projects. This paper discusses the ongoing effort to digitize and catalog TSTA’s historical data to create a searchable resource for the fusion research community. While the long-term objective is to develop a relational database for structured data management, the project remains in the early phase, with current efforts focused on scanning and indexing physical documents. Initial plans for database implementations are also presented, outlining key considerations for structure, query indexing, and standardization. As digitization progresses, future discussions will refine these implantation details to ensure an efficient and comprehensive system. This initiative aims to preserve critical legacy data, enhance the design of tritium system facilities, and support the next generation of fusion energy research.

42 ENGINEERING↗

A High Performance Sparse Tensor Algebra Compiler in MLIR

Sparse tensor algebra is widely used in many applications, including scientific computing, machine learning, and data analytics. The performance of sparse tensor algebra kernels strongly depends on the intrinsic characteristics of the input tensors, hence many storage formats are designed for tensors to achieve optimal performance for particular applications/architectures, which makes it challenging to implement and optimize every tensor operation of interest on a given architecture. We propose a tensor algebra domain-specific language (DSL) and compiler framework to automatically generate kernels for mixed sparse-dense tensor algebra operations. The proposed DSL provides high-level programming abstractions that resemble the familiar Einstein notation to represent tensor algebra operations. The compiler introduces a new Sparse Tensor Algebra dialect built on top of LLVM's extensible MLIR compiler infrastructure for efficient code generation while covering a wide range of tensor storage formats. Our compiler also leverages input-dependent code optimization to enhance data locality for better performance. Our results show that the performance of automatically generated kernels outperforms the state-of-the-art sparse tensor algebra compiler, with up to 20.92x, 6.39x, and 13.9x performance improvement over state-of-the-art tensor algebra compilers, for parallel SpMV, SpMM, and TTM, respectively.

Tian, Ruiqin↗

Learning to Correct Climate Projection Biases

The fidelity of climate projections is often undermined by biases in climate models due to their simplification or misrepresentation of unresolved climate processes. While various bias correction methods have been developed to post-process model outputs to match observations, existing approaches usually focus on limited, low-order statistics, or break either the spatiotemporal consistency of the target variable, or its dependency upon model resolved dynamics. We develop a Regularized Adversarial Domain Adaptation (RADA) methodology to overcome these deficiencies, and enhance efficient identification and correction of climate model biases. Instead of pre-assuming the spatiotemporal characteristics of model biases, we apply discriminative neural networks to distinguish historical climate simulation samples and observation samples. The evidences based on which the discriminative neural networks make distinctions are applied to train the domain adaptation neural networks to bias correct climate simulations. We regularize the domain adaptation neural networks using cycle-consistent statistical and dynamical constraints. An application to daily precipitation projection over the contiguous United States shows that our methodology can correct all the considered moments of daily precipitation at approximately $1^\circ$ resolution, ensures spatiotemporal consistency and inter-field correlations, and can discriminate between different dynamical conditions. Our methodology offers a powerful tool for disentangling model parameterization biases from their interactions with the chaotic evolution of climate dynamics, opening a novel avenue toward big-data enhanced climate predictions.

58 GEOSCIENCES↗

DriveSense: A Noise-Resilient Framework for Driving Mode Identification

Accurate drive mode classification is essential for enhancing the reliability and predictive maintenance of heavy-duty electric trucks. This study proposes a novel fuzzy logic-based framework, DriveSense, for real-time drive mode classification, addressing key challenges such as sensor noise, transitional behaviors, and computational efficiency. The proposed approach integrates a two-stage filtering pipeline, combining adaptive outlier removal and a dynamic Kalman filter to enhance data quality. A fuzzy inference system with smoothened trapezoidal membership functions is then applied to classify driving modes into standstill, constant speed, acceleration, and deceleration while mitigating the effects of noise and edge cases. Performance evaluation using real-world and simulated drive cycles demonstrates significant improvements in classification accuracy (up to 97.8%), F1-score (up to 0.97), and robustness against noise, while reducing false positives. Comparative analysis against baseline models, demonstrates DriveSense’s superior accuracy and generalizability across diverse driving patterns. The framework’s lightweight and interpretable fuzzy inference engine operates with low computational latency, ensuring compatibility with real-time embedded systems typical of heavy-duty electric trucks. Moreover, DriveSense models transitional behaviors through overlapping fuzzy sets and adaptive borderline classification logic, enabling smooth identification of subtle shifts such as rolling stops or gradual deceleration. These results highlight DriveSense’s potential to enhance predictive maintenance strategies, reduce downtime, and support scalable, fleet-wide diagnostics.

Kumar, Praveen [Oak Ridge National Laboratory (ORN↗

Machine Learning for Optical Scanning Probe Nanoscopy

Abstract The ability to perform nanometer‐scale optical imaging and spectroscopy is key to deciphering the low‐energy effects in quantum materials, as well as vibrational fingerprints in planetary and extraterrestrial particles, catalytic substances, and aqueous biological samples. These tasks can be accomplished by the scattering‐type scanning near‐field optical microscopy (s‐SNOM) technique that has recently spread to many research fields and enabled notable discoveries. Herein, it is shown that the s‐SNOM, together with scanning probe research in general, can benefit in many ways from artificial‐intelligence (AI) and machine‐learning (ML) algorithms. Augmented with AI‐ and ML‐enhanced data acquisition and analysis, scanning probe optical nanoscopy is poised to become more efficient, accurate, and intelligent.

Chen, Xinzhong↗

Unveiling the nanoscale architectures and dynamics of protein assembly with in situ atomic force microscopy

Proteins play a vital role in different biological processes by forming complexes through precise folding with exclusive inter- and intra-molecular interactions. Understanding the structural and regulatory mechanisms underlying protein complex formation provides insights into biophysical processes. Furthermore, the principle of protein assembly gives guidelines for new biomimetic materials with potential applications in medicine, energy, and nanotechnology. Atomic force microscopy (AFM) is a powerful tool for investigating protein assembly and interactions across spatial scales (single molecules to cells) and temporal scales (milliseconds to days). It has significantly contributed to understanding nanoscale architectures, inter- and intra-molecular interactions, and regulatory elements that determine protein structures, assemblies, and functions. This review describes recent advancements in elucidating protein assemblies with in situ AFM. We discuss the structures, diffusions, interactions, and assembly dynamics of proteins captured by conventional and high-speed AFM in near-native environments and recent AFM developments in the multimodal high-resolution imaging, bimodal imaging, live cell imaging, and machine-learning-enhanced data analysis. These approaches show the significance of broadening the horizons of AFM and enable unprecedented explorations of protein assembly for biomaterial design and biomedical research.

36 MATERIALS SCIENCE↗

Empirical validation of building energy simulation model input parameter for multizone commercial building during the cooling season

This paper presents a critical advancement in Building Energy Modeling (BEM) through an empirical validation approach using a high-quality dataset from a multizone commercial office building in Oak Ridge, TN, USA. BEM is widely utilized in diverse construction applications, but its effectiveness relies on the accuracy of its predictions. The study focuses on empirical validation of input parameters in BEM, including building envelope data, infiltration modeling, and rooftop unit system performance curves. The validation of simulation input parameters leads to substantial improvements in the accuracy of simulation results. Notable both NMBE and cv (RMSE) values are reduced by 0.5 % for indoor air temperature and 17 % for indoor air relative humidity compared to the previous model. At the system level, both NMBE and cv (RMSE) values are reduced by 2 % for fan energy consumption and 4 % for cooling energy consumption, compared to the previous model. A literature review highlights a significant gap in empirical validation studies, which predominantly concentrate on either component-level or whole building validation. Furthermore, many studies employ simplified setups that may not faithfully represent the complexities of multizone commercial buildings. This paper distinguishes itself by emphasizing the critical importance of component-level input parameter validation. It underlines the need to validate data related to building envelope components and HVAC system performance curves, resulting in more accurate simulation outcomes. In conclusion, the utilization of actual multizone commercial building data enhances the study's practical relevance. In summary, this research underscores the pivotal role of input parameter validation in enhancing the accuracy and reliability of BEM.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Aggregation Methods for Quantifying PTM and Structural Changes in Bottom-Up Proteomics

Bottom-up proteomic workflows rely on sequential preprocessing steps, commonly including peptide-to-protein aggregation (“roll-up”), to enhance data reliability and interpretability. While roll-up is effective for protein-centered analyses, it may be suboptimal for applications focused on post-translational modifications (PTMs) or protein structural changes, such as limited proteolysis–mass spectrometry (LiP-MS). Here, we investigate how different roll-up strategies influence site-level quantification in PTM differential analysis. Moreover, we introduce a novel site-centric roll-up approach tailored for LiP-MS, which quantifies proteolytic fragments rather than solely tryptic peptides. We benchmark these methods through simulation studies, comparing their sensitivity and specificity in detecting structural and PTM-driven changes. We found that the median and mean roll-up methods outperform the sum method in both PTM and LiP proteomics, and site-level quantification in LiP outperforms peptide-level quantification. Our findings offer the first systematic, data-driven guidance for selecting roll-up techniques in site-level proteomic analyses, with implications for both PTM-focused and structural proteomics studies.

aggregation↗

Expanded Understanding of the Western Antarctic Peninsula Sea‐Ice Environment Through Local and Regional Observations at Palmer Station

Abstract The Western Antarctic Peninsula (WAP) has been experiencing rapid regional warming since at least the 1950s, however, the impacts of this warming at the local scale are variable and nuanced. Previous studies that have linked sea‐ice variability to biogeochemical cycles and food web dynamics often combine local‐scale biogeochemical data with coarse‐resolution regional satellite sea‐ice data, which may not adequately capture local sea‐ice conditions. In this study, we analyzed local‐scale in situ sea‐ice observations collected as part of a 28‐year record (1992–2020) from the Palmer Long‐Term Ecological Research site at Anvers Island, mid‐WAP, in conjunction with isotopically‐derived sea‐ice meltwater (SIM) fractions and satellite‐derived sea‐ice motion and concentration, to quantify the variability and long‐term trends in local sea‐ice behavior. In situ sea ice observations at Palmer Station displayed higher variability than satellite observations and showed no significant declines over this time, despite region‐wide declines identified in prior studies. Higher spring SIM fractions were attributed to strong northward sea‐ice motion throughout the winter. Applying these local‐scale sea‐ice insights to similarly scaled stratification and chlorophyll‐ a measurements, we found that a longer‐lasting, more consistent sea‐ice pack led to greater water column stratification following the spring sea‐ice retreat. Greater sea‐ice persistence and stronger stratification led to larger peaks in chlorophyll‐ a , though sea‐ice metrics did not explain the positive temporal trends in either stratification strength or chlorophyll‐ a . Through this study, we identify how local sea‐ice observations and meltwater data can enhance satellite data to build an understanding of the intricate connections between ice, water column dynamics, and phytoplankton.

Goodell, E.↗

Remote Sensing of Live Fuel Moisture for Wildfires Using SMAP Satellite Observations

Live Fuel Moisture (LFM) is a critical parameter for wildfire risk assessment, traditionally measured by labor-intensive field sampling. However, sampled LFM data are influenced by site-specific factors, such as local vegetation types and plant traits, and are often collected retrospectively after wildfire events, making it difficult to obtain pre-fire data for predictive applications. Here, we evaluate the relationship between LFM and Vegetation Water Content (VWC) and Soil Moisture (SM) retrieved from SMAP L-band brightness temperature using the Maximum Entropy Production (MEP) approach. The MEP-retrieved VWC exhibited strong correlation with in situ measurements of LFM ( r > 0.6) in the Western U.S. The integration of high-resolution vegetation coverage data enhances the detection of sub-grid vegetation heterogeneity. This study demonstrates the operational potential of remote sensing derived VWC as a scalable proxy of LFM, supporting its application in regional assessment of wildfire risk.

Cho, Kyeungwoo [Georgia Institute of Technology, A↗