Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “science data analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Visualization Within the Department of Energy: NREL IEEE VIS Application Spotlight

This presentation highlights the role of advanced visualization techniques at the National Renewable Energy Laboratory (NREL) in supporting cutting-edge research across diverse energy domains. From immersive analytics and uncertainty visualization to high-resolution and real-time data analysis, NREL's visualization capabilities enable scientists to explore complex datasets more effectively. These tools are critical for advancing research in materials science, renewable energy technologies, biofuels, electric vehicle infrastructure, energy efficiency - from industrial processes to entire communities - and then bringing these innovations to practice through energy systems integration. NREL's visualization tools drive innovation across renewable energy and grid modernization efforts by providing deeper insights and improving decision-making.

grid modernization↗

RUBISCO Soil Moisture Working Group (SMWG) Mini-Workshop Report

The RUBISCO Soil Moisture Working Group (SMWG), an initiative bolstered by the DOE RUBISCO Science Focus Area project, is committed to unraveling the intricacies of soil moisture and its profound impact on global hydrology, energy, and biogeochemical cycles. Confronted with the complex challenges in multi-scale SM data development, feedback analysis, and model benchmarking, the SMWG aims to catalyze advancements in the field through a synergistic approach that engages leaders in soil moisture research and encompasses diverse datasets and novel analytical techniques.

54 ENVIRONMENTAL SCIENCES↗

Aligning NASA Earth Science Data Stewardship with FAIR Principles: Outcomes, Recommendations, and Future Directions

The FAIR Principles—Findable, Accessible, Interoperable, and Reusable—offer a widely accepted framework for improving the sharing and reuse of digital scientific data by both human and machine users. Following these principles is critical for effective scientific data stewardship, broader scientific collaboration, and compliance with federal and agency data policies. This paper, based on the work of NASA’s Open, Free, and FAIR Working Group (O’FAIR WG) under the Earth Science Data Systems Program, presents an overview of how FAIR is being applied within NASA’s Earth science data landscape. It highlights ongoing progress and challenges, identifies FAIR-enabling resources, and offers recommendations and strategic actions to enhance the FAIRness of NASA-funded open and free Earth science data products. The FAIR-enabling resources identified underscore the vital role of NASA's existing enterprise processes, standards, tools, and infrastructures in supporting FAIR implementation. Our findings show strong performance in making NASA Earth science data more findable and accessible. However, further work is needed—especially in enhancing interoperability, so that different systems and tools can better understand and exchange data. This is especially important for enabling machine-driven discovery and analysis. We emphasize the importance of a balanced strategy that combines a centralized, top-down approach—focused on building enterprise-level capabilities and processes—with a decentralized, bottom-up approach driven by discipline-specific needs and community practices. We advocate for coordinated efforts to enhance (meta)data interoperability to facilitate seamless data and information sharing and exchange of Earth science data both within NASA and across other agencies managing Earth science data.

Data Product↗

Size-resolved particle and black carbon deposition over the cryosphere (Final Report)

Our project focused on providing observational constraints and investigation of aerosol deposition by performing the first unambiguous direct eddy flux covariance measurements of aerosol dry deposition over the cryosphere. The project aimed to make eddy covariance flux measurements of dry deposition and off-line wet deposition measurements of black carbon containing aerosol plus eddy covariance measurements of size-resolved accumulation mode scattering aerosol over the DOE ARM AMF3 Oliktok Point site in Fall 2020. We aimed to compare our measurements to model parameterizations to constrain uncertainties and systematic bias, and improve deposition parameterizations in models. Due to disruptions in field work from COVID-19, our work was delayed and moved to an alternate site on the North Slope of Alaska. This report summarizes the key findings from the project, including data analysis from previous field projects and from the North Slope of Alaska.

54 ENVIRONMENTAL SCIENCES↗

A comparative analysis of YOLOv8 and U-Net image segmentation approaches for transmission electron micrographs of polycrystalline thin films

Metallic thin films offer a platform to experimentally study the dynamics of microstructural evolution, but the required transmission electron microscopy (TEM)-based imaging generates complex images that are challenging to segment and quantify. This work provides a comparative analysis of a new YOLOv8 model and an established U-Net model for bright-field TEM images of polycrystals, employing a framework leveraging physical observables to evaluate performance against two hand-traced benchmark datasets. This methodology obviates the comparison of large, diversely structured, and manually labeled datasets that are required to assess performance on a per-image/per-pixel basis. It is found that the YOLOv8 model, adapted for real-time instance segmentation, has up to 43× faster inferencing (NVIDIA GeForce RTX 4090) compared to U-Net and reconstructs hand-traced grain size distributions (GSDs) with excellent fidelity, finding mean diameter within 3% for grains near an optimal magnification; for grains that deviate from the optimal pixel-diameter, the size of small- (large)-diameter grains is systematically over- (under)-estimated. This is partially mitigated by including scale-aware augmentations during training. Moreover, when the bias is corrected post-inference by a rigid shift in distribution, the YOLOv8 model reproduces ground truth GSDs with exceptional fidelity, with statistical tests indicating <5% probability that the distributions are distinct. Based on ground truth data, calibration curves pertaining to this shift can be constructed for a given model. This issue is not present in the U-Net model’s results, indicating that for quantitative measurements where the true size of objects is of interest, special procedures must be implemented for YOLO-based models.

36 MATERIALS SCIENCE↗

Digital Twin for Chemical Science (DTCS) v0.01

Directly visualizing the trajectories of chemistry can unravel novel insights into the behavior of catalysts, gas phase reactions, photo-induced dynamics, and building blocks for quantum information processing. The ability of explicitly identifying, tracking, and tagging the exchange of matter, hence the annihilation and creation of new chemical species, can be best realized through a close coupling of theory and experiment. While the synchrotron-based characterization facilities propelled rapidly in its hardware, providing higher brightness, better resolution, and more precision, the software infrastructure is lagging. We developed DTCS (Digital Twin for Chemical Science) v.01, a central platform that faithfully mimics advanced instrumentations in Scientific User Facilities, by solving a variety of technical challenges in data acquisition, analysis, and model-driven interpretation. Rooted in physics and accelerated by AI, we validated this concept by direct comparison with precise experimental X-ray Photoelectron Spectroscopy (XPS) observations using a ubiquitous metal-water interfacial scenario, i.e., Ag/H2O as our main narrative. The DTCS v.01 input mirrors how the bench chemists work, with the output directly linked to the end station computer, thereby providing a user-friendly, knowledge-driven, and accessible user experience with mechanistic insights standardized in a way that are ready to be published, versioned, and transferred flexibly.

Qian, Jin↗

High-count-rate effects in event processing for the XRISM/Resolve X-ray microcalorimeter. II. Energy scale and resolution in orbit

The Resolve instrument on the X-ray Imaging and Spectroscopy Mission (XRISM) uses a 36 pixel microcalorimeter designed to deliver high-resolution, non-dispersive X-ray spectroscopy. Although it is optimized for extended sources with low count rates, Resolve observations of bright point sources are still able to provide unique insights into the physics of these objects, as long as high-count-rate effects are addressed in the analysis. These effects include the loss of exposure time for each pixel, changes in the energy scale, and changes in the energy resolution. To investigate these effects under realistic observational conditions, we observed the bright X-ray source, the Crab Nebula, with XRISM at several offset positions with respect to the Resolve field of view and with continuous illumination from 55 Fe sources on the filter wheel. For the spectral analysis, we excluded data where exposure-time loss was too significant to ensure reliable spectral statistics. The energy scale at 6 keV shows a slight negative shift in the high-count-rate regime. The energy resolution at 6 keV worsens as the count rate in electrically neighboring pixels increases, but can be restored by applying a nearest-neighbor coincidence cut (“cross-talk cut”). We examined how these effects influence the observation of bright point sources, using GX 13+1 as a test case, and identified an eV-scale energy offset at 6 keV between the inner (brighter) and outer (fainter) pixels. Users who seek to analyze velocity structures on the order of tens of km s–1 should account for such high-count-rate effects. These findings will aid in the interpretation of Resolve data from bright sources and provide valuable considerations for designing and planning for future microcalorimeter missions.

X-rays: general↗

Survey of Deep Learning and Physics-Based Approaches in Computational Wave Imaging

Computational wave imaging (CWI) extracts hidden structure and physical properties of a volume of material by analyzing wave signals that traverse that volume. Applications include seismic exploration of the Earth’s subsurface, acoustic imaging and nondestructive testing (NDT) in material science, and ultrasound computed tomography (USCT) in medicine. Current approaches for solving CWI problems can be divided into two categories: those rooted in traditional physics and those based on deep learning. Physics-based methods stand out for their ability to provide high-resolution and quantitatively accurate estimates of acoustic properties within the medium. However, they can be computationally intensive and are susceptible to ill-posedness and nonconvexity typical of CWI problems. Machine learning (ML)-based computational methods have recently emerged, offering a different perspective to address these challenges. Diverse scientific communities have independently pursued the integration of deep learning in CWI. This review discusses how contemporary scientific ML techniques, and deep neural networks in particular, have been developed to enhance and integrate with traditional physics-based methods for solving CWI problems. We present a structured framework that consolidates existing research spanning multiple domains, including computational imaging, wave physics, and data science. This study concludes with important lessons learned from existing ML-based methods and identifies technical hurdles and emerging trends through a systematic analysis of the extensive literature on this topic.

42 ENGINEERING↗

Integrated Methane Monitoring Platform Extension, Volume I: Final Technical Report

The IMMPE project, DE-FE0032284, was to enhance methane monitoring technologies and their applications across various natural gas asset classes. The scope included deploying advanced methane detection and monitoring technologies to identify and mitigate fugitive methane emissions, measuring emission rates, and assessing impacts. The findings included the successful mitigation of identified emissions and quantification of emission rates. A key outcome was the development of a comprehensive template and summary of recommendations for methane emissions monitoring, which is replicable for both upstream and downstream applications. Furthermore, the project emphasized the importance of education by providing training opportunities for technicians and regulators, thereby fostering awareness and promoting the adoption of cost-effective methane emissions monitoring and management techniques.

02 PETROLEUM↗

A data-driven framework for predicting machining stability: employing simulated data, operational modal analysis, and enhanced transfer learning

Chatter, a self-excited vibration phenomenon, presents a significant challenge in machining operations, particularly in high-speed milling, where it can degrade tool life, reduce material removal efficiency, and compromise workpiece quality. Addressing this challenge requires a reliable predictive model that can accommodate the complex dynamics of various machining scenarios. This study introduces a novel, data-driven approach to predicting machining stability, leveraging over 140,000 simulated datasets and employing advanced techniques such as operational modal analysis (OMA), enhanced transfer learning (TL), and receptance coupling substructure analysis (RCSA). By integrating these methodologies, the framework effectively classifies and predicts chatter across diverse operational modes, achieving robust and accurate outcomes. Our model utilizes a Random Forest (RF) classifier trained with the comprehensive dataset, which demonstrates substantial improvements in both predictive accuracy and robustness. Specifically, the RF model achieved an accuracy rate of 85%, an area under the curve (AUC) of 0.90, and an F1 score of 0.88, underscoring its capability to adapt to varying machining configurations. These results highlight the framework’s potential to enhance operational efficiency and machining quality by providing reliable chatter predictions across a broad range of machining parameters. In conclusion, this research thus offers a significant advancement in predictive maintenance for machining processes, enabling more stable and efficient manufacturing operations.

42 ENGINEERING↗

Source Levels of In‐Cloud Air in Shallow Cumulus: Consistency Between Paluch Diagram and Lagrangian Particle Tracking

Abstract The Paluch diagram is a widely used tool for interpreting aircraft measurements of shallow cumulus clouds. A prior study conducted by Heus et al. (2008,https://doi.org/10.1175/2008jas2572.1) concluded that the source levels of in‐cloud air inferred from the Paluch diagram exhibit biases, sometimes of several hundred meters, in comparison to those derived from Lagrangian particle tracking. In this short study we revisit this comparison. The results indicate that the upper source levels of in‐cloud air determined from the Lagrangian Particle Tracking and the Paluch diagram are consistent, and the choice of statistical methods is crucial. The significance of this research lies in confirming the reliability of the Paluch analysis, enabling its confident application to aircraft data.

Meteorology & Atmospheric Sciences↗

Semi–Analytical Modeling of Transient Stream Drawdown and Depletion in Response to Aquifer Pumping

Analytical and semi–analytical models for stream depletion with transient stream stage drawdown induced by groundwater pumping are developed to address a deficiency in existing models, namely, the use of a fixed stream stage condition at the stream–aquifer interface. Here field data are presented to demonstrate that stream stage drawdown does indeed occur in response to groundwater pumping near aquifer–connected streams. A model that predicts stream depletion with transient stream drawdown is developed based on stream channel mass conservation and finite stream channel storage. The resulting models are shown to reduce to existing fixed–stage models in the limit as stream channel storage becomes infinitely large, and to the confined aquifer flow with a no–flow boundary at the streambed in the limit as stream storage becomes vanishingly small. The model is applied to field measurements of aquifer and stream drawdown, giving estimates of aquifer hydraulic parameters, streambed conductance, and a measure of stream channel storage. The results of the modeling and data analysis presented herein have implications for sustainable groundwater management.

54 ENVIRONMENTAL SCIENCES↗

Data for “Tree root nutrient uptake kinetics vary with nutrient availability, environmental conditions, and root traits: A global analysis”

This data package contains data and code used in the paper “Tree root nutrient uptake kinetics vary with nutrient availability, environmental conditions, and root traits: A global analysis”. The central product is a global dataset of root inorganic nutrient uptake rates and kinetics parameters covering temperate, boreal, and sub/tropical tree species, representing a collection of nutrient uptake data from published studies. This dataset enables tree investigation of root nutrient uptake rates across species, space, and experimental conditions. The data can also be combined with supplementary data on root and soil traits or with external datasets (e.g. R scripts contained within use data from FRED 3.0; (Iversen et al., 2021)). Contained within is the main nutrient data “uptake_data.csv” as well as 4 additional .csv files that link uptake data to supplementary measurements, source references, taxonomic information, and additional nutrient uptake measurements across nutrient gradients, and 1 .csv file that records meta-analysis results for plotting with the R scripts. There are seven R scripts that support data analysis and creation of the figures in the related publication.

54 ENVIRONMENTAL SCIENCES↗

NDMAS

Overview of Current ART-GCR Data: Fuel Fabrication, Irradiation Monitoring (Fuel & Graphite – near real-time for HDG-1), Post-Irradiation Examination (Fuel & Graphite), Graphite Characterization (Baseline and Irradiated), High Temperature Metals Mechanical Tests, Design, Methods, and Validation Data, Japan Atomic Energy Agency’s High Temperature Test Reactor (HTTR), Argonne National Laboratory’s Natural convection Shutdown heat removal Test Facility (NSTF), Oregon State University’s High Temperature Test Facility (HTTF), Generation IV International VHTR Materials Handbook, Additional related data, and Advanced Test Reactor operations (near real-time).

11 - NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

OLCF’s Advanced Computing Ecosystem (ACE): FY25 Update for Ongoing Efforts

The advent of widespread use of artificial intelligence (AI) and machine learning (ML) models in science, coupled with fast data production rates of scientific instruments strain the traditional batch-oriented high-performance computing (HPC) environment. As scientific exploration continues to require more data and faster processing and analysis, new emerging technologies and capabilities to enable cross-facility and time-sensitive workflows are required for seamless integration of HPC and experimental facilities. The Advanced Computing Ecosystem (ACE) is a strategic initiative within the Oak Ridge Leadership Computing Facility (OLCF) established in 2024 to support the development of cutting-edge technologies to advance computational research and infrastructure at OLCF and across the Department of Energy (DOE). Several DOE initiatives are spearheading the evolution of the scientific landscape by blurring facility boundaries and connecting the user facilities to advance scientific capabilities and ensure energy dominance. The DOE Integrated Research Infrastructure (IRI) program is one example that is laying a foundation to support complex cross-facility workflows. The IRI program aims to integrate diverse computational resources, data infrastructures, and scientific instruments to facilitate collaboration and accelerate scientific discovery. The Interconnected Science Ecosystem (INTERSECT) initiative at Oak Ridge National Laboratory (ORNL) is another example that aims to revolutionize scientific research through AI-driven, interconnected autonomous laboratories and research facilities. Finally, the American Science Cloud (AmSC), recently announced in the “One Big Beautiful Bill”, aims to leverage prior infrastructure efforts of the IRI and automation and AI efforts of INTERSECT (and others) to build a federated, AI-augmented AmSC platform to unify the DOE’s computing, experimental, and data resources to catalyze scientific innovation.

97 MATHEMATICS AND COMPUTING↗

MSD CoP Webinar: "Advances in MSD-LIVE to Support the MSD Community of Practice"

Context: This webinar was hosted by the MultiSector Dynamics Community of Practice (MSD CoP; https://multisectordynamics.org). Advances in MSD-LIVE to Support the MSD Community of Practice Presenters: Casey Burleyson and Zoe Guillen (Pacific Northwest National Laboratory) Abstract: The MultiSector Dynamics Living, Intuitive, Value-adding, Environment (MSD-LIVE; msdlive.org) is a cloud-based data management system and advanced computing platform that enables MSD researchers to document and archive their data, run their models and analysis tools, and share their data, software, and workflows within the MSD Community of Practice. Recently, several high-profile datasets have attracted many new users to MSD-LIVE. This webinar has two goals: 1) To refamiliarize the MSD community and new users with the components of the platform (e.g., the data repository, model training notebooks, and data dashboards) and to highlight examples of how these components are advancing MSD science and 2) To demonstrate new features in v3 of the platform, released in late 2025. The main new feature in v3 is the ability to interactively explore data in MSD-LIVE without downloading it. MSD-LIVE users can now click a button in our data repository and launch a blank Jupyter notebook with access to the underlying data on AWS. Users can use the notebook to write analysis, visualization, or subsetting routines that process the data directly on the AWS cloud. We also added a GitHub integration feature that allows users to share analysis or visualization code they develop with the community of MSD-LIVE users. The webinar will wrap up with a look at what's coming next for MSD-LIVE in 2026. Moderator: Patrick M. Reed (MSD CoP Facilitation Team) This webinar was held on: May 12th, 2026 from 1-2 PM EST.

Open Science↗

A Survey on the Expanding Scope and Interdisciplinary Opportunities for Processing-in-Memory Techniques

Processing-in-Memory (PIM) is emerging as a practical path to overcome the limitations of traditional von Neumann architectures. At its core, PIM systems implement computing primitives such as logic operations and multiply-accumulate acceleration through compute-in-memory, near-memory processing, or hybrid designs. The role of memory cells varies widely across technologies, acting as inputs, outputs, or analog accumulators through bit-lines and sense amplifiers. This diversity creates trade-offs in precision, bandwidth, latency, and programmability, making it difficult to build a unified understanding on the progress of the field. In this survey, we organize recent advances of PIM into three areas. First, we discuss the progress on the architectural optimizations of PIM and its integration with both DRAM and emerging non-volatile memories. Second, we examine how PIM is being used to accelerate key computing domains, including generative AI workloads and high-performance kernels, along with new approaches. Third, we highlight the growing adoption of PIM in computational sciences, where it is being applied to solve interdisciplinary problems such as genome analysis, mRNA quantification, mass spectrometry, quantum circuit simulation, wave modeling, and secure computation. Finally, we synthesize the major challenges that continue to slow PIM adoption, including manufacturing constraints, power delivery, thermal reliability, data consistency, runtime and memory-management coordination, and the difficulty of building portable software abstractions without sacrificing commercial viability. This work provides an updated, structured perspective on PIM’s potential across computing and computational sciences and the barriers that must be solved for it to reach its full impact.

Asifuzzaman, Kazi [Oak Ridge National Laboratory (↗