Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “standardized data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

Multi-Site Observational Study to Assess Biomarkers for Susceptibility or Resilience to Chronic Pain: The Acute to Chronic Pain Signatures (A2CPS) Study Protocol

Chronic pain has become a global health problem contributing to years lived with disability and reduced quality of life. Advances in the clinical management of chronic pain have been limited due to incomplete understanding of the multiple risk factors and molecular mechanisms that contribute to the development of chronic pain. The Acute to Chronic Pain Signatures (A2CPS) Program aims to characterize the predictive nature of biomarkers (brain imaging, high-throughput molecular screening techniques, or “omics,” quantitative sensory testing, patient-reported outcome assessments and functional assessments) to identify individuals who will develop chronic pain following surgical intervention. The A2CPS is a multisite observational study investigating biomarkers and collective biosignatures (a combination of several individual biomarkers) that predict susceptibility or resilience to the development of chronic pain following knee arthroplasty and thoracic surgery. This manuscript provides an overview of data collection methods and procedures designed to standardize data collection across multiple clinical sites and institutions. Pain-related biomarkers are evaluated before surgery and up to 3 months after surgery for use as predictors of patient reported outcomes 6 months after surgery. The dataset from this prospective observational study will be available for researchers internal and external to the A2CPS Consortium to advance understanding of the transition from acute to chronic postsurgical pain.

60 APPLIED LIFE SCIENCES↗

Steam Condensation Scaled Experiment in the Presence of Non-condensable Gas for Reactor Containment Passive Safety Analysis

This study presents scaled experiments using steam condensation with non-condensable gas (NCG)—helium, simulating hydrogen—as these experiments are pivotal for water-cooled reactor passive containment cooling system (PCCS) design and analysis. Research into PCCSs for small modular reactors (SMRs) is especially important in light of SMR system design; however, studies in the literature reflect limitations due to test geometry and operational condition variations, without considering SMR prototypic design. To address these challenges, a scaled test facility was developed to accurately replicate SMR PCCSs. This facility includes vertical down-flow condensing test sections with 1-, 2-, and 4-in.-diameter condensing tubes, accompanied by annular water cooling. Experiments were conducted using both superheated and saturated steam, with steam mass flow rates varying from 55 to 66 kg/hr., in the presence of helium as the NCG mass flow rate ranges from 1.8 to 22 kg/hr. Test data were collected on (a) the axial temperatures of the annular cooling water; (b) the outer wall temperature of the condensers; and (c) the mass flow rate, temperature, and pressure at the test section inlets and outlets. These primary test data were used in conjunction with a standard data reduction methodology to estimate essential thermal parameters such as heat fluxes, heat transfer coefficients, and condensation rates. The effects of NCGs on steam condensation within the geometry of the scaled test sections were then presented in regard to various testing conditions.

21 SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLANTS↗

Presentation: Steam Condensation Scaled Experiment in the Presence of Non-condensable Gas for Reactor Containment Passive Safety Analysis

This study presents scaled experiments using steam condensation with non-condensable gas (NCG)—helium, simulating hydrogen—as these experiments are pivotal for water-cooled reactor passive containment cooling system (PCCS) design and analysis. Research into PCCSs for small modular reactors (SMRs) is especially important in light of SMR system design; however, studies in the literature reflect limitations due to test geometry and operational condition variations, without considering SMR prototypic design. To address these challenges, a scaled test facility was developed to accurately replicate SMR PCCSs. This facility includes vertical down-flow condensing test sections with 1-, 2-, and 4-in.-diameter condensing tubes, accompanied by annular water cooling. Experiments were conducted using both superheated and saturated steam, with steam mass flow rates varying from 55 to 66 kg/hr., in the presence of helium as the NCG mass flow rate ranges from 1.8 to 22 kg/hr. Test data were collected on (a) the axial temperatures of the annular cooling water; (b) the outer wall temperature of the condensers; and (c) the mass flow rate, temperature, and pressure at the test section inlets and outlets. These primary test data were used in conjunction with a standard data reduction methodology to estimate essential thermal parameters such as heat fluxes, heat transfer coefficients, and condensation rates. The effects of NCGs on steam condensation within the geometry of the scaled test sections were then presented in regard to various testing conditions.

21 SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLANTS↗

A Comprehensive Northern Hemisphere Particle Microphysics Data Set From the Precipitation Imaging Package

Microphysical observations of precipitating particles are critical data sources for numerical weather prediction models and remote sensing retrieval algorithms. However, obtaining coherent data sets of particle microphysics is challenging as they are often unindexed, distributed across disparate institutions, and have not undergone a uniform quality control process. This work introduces a unified, comprehensive Northern Hemisphere particle microphysical data set from the National Aeronautics and Space Administration precipitation imaging package (PIP), accessible in a standardized data format and stored in a centralized, public repository. Data is collected from 10 measurement sites spanning 34° latitude (37°N–71°N) over 10 years (2014–2023), which comprise a set of 1,070,000 precipitating minutes. The provided data set includes measurements of a suite of microphysical attributes for both rain and snow, including distributions of particle size, vertical velocity, and effective density, along with higher-order products including an approximation of volume-weighted equivalent particle densities, liquid equivalent snowfall, and rainfall rate estimates. The data underwent a rigorous standardization and quality assurance process to filter out erroneous observations to produce a self-describing, scalable, and achievable data set. Case study analyses demonstrate the capabilities of the data set in identifying physical processes like precipitation phase-changes at high temporal resolution. Bulk precipitation characteristics from a multi-site intercomparison also highlight distinct microphysical properties unique to each location. This curated PIP data set is a robust database of high-quality particle microphysical observations for constraining future precipitation retrieval algorithms, and offers new insights toward better understanding regional and seasonal differences in bulk precipitation characteristics.

54 ENVIRONMENTAL SCIENCES↗

Microbiome data management in action workshop: Atlanta, GA, USA, June 12–13, 2024

Microbiome research is revolutionizing human and environmental health, but the value and reuse of microbiome data are significantly hampered by the limited development and adoption of data standards. While several ongoing efforts are aimed at improving microbiome data management, significant gaps still remain in terms of defining and promoting adoption of consensus standards for these datasets. The Strengthening the Organization and Reporting of Microbiome Studies (STORMS) guidelines for human microbiome research have been endorsed and successfully utilized by many research organizations, publishers, and funding agencies, and have been recognized as a consensus community standard. No equivalent effort has occurred for environmental, synthetic, and non-human host-associated microbiomes. To address this growing need within the microbiome research community, we convened the Microbiome Data Management in Action Workshop (June 12–13, 2024, in Atlanta, GA, USA), to bring together key decision makers in microbiome science including researchers, publishers, funders, and data repositories. The 50 attendees, representing the diverse and interdisciplinary nature of microbiome research, discussed recent progress and challenges, and brainstormed actionable recommendations and paths forward for coordinated environmental microbiome data management and the modifications necessary for the STORMS guidelines to be applied to environmental, non-human host, and synthetic microbiomes. The outcomes of this workshop will form the basis of a formalized data management roadmap to be implemented across the field. These best practices will drive scientific innovation now and in years to come as these data continue to be used not only in targeted reanalyses but in large-scale models and machine learning efforts.

54 ENVIRONMENTAL SCIENCES↗

Historic Source Inspection Metadata

Although this data dictionary describes the raw data format, it is relevant only for data entered in 2017 and later. Entries in this date range represent formatted, standardized data that have been verified against original, raw data files by the database manager. She took ownership of the database in September 2022 and was able to standardize entries by comparing against raw data files back to 2017. Furthermore, for comparison of historic pre-source inspection data to new pre-source data supplied by the vendor and intended for action limit analysis purposes, this time period is of sufficient length.

97 MATHEMATICS AND COMPUTING↗

From Silos to Synergy: Identifying a Roadmap for Cross-Sector Research to Accelerate the Clean Energy Transition

The U.S. Department of Energy's blueprints for the transportation, buildings, and electricity sectors call for substantial reductions in greenhouse gas (GHG) emissions by 2050. These plans focus on zero-emission vehicles, investments in transit, energy-efficient buildings, and the widespread adoption and deployment of renewable energy technologies like solar photovoltaics (PV), energy storage and energy-efficient appliances. However, these sectors are often studied and modeled in isolation, overlooking how household decisions to adopt clean technologies in one sector influence others. This study, led by an interdisciplinary team at the National Renewable Energy Laboratory (NREL), explores opportunities for cross-sector collaboration to drive more effective and equitable decarbonization. Through discussions with 22 NREL researchers across transportation, building, solar, and grid sectors, the study highlights the need for integrated tools and models that capture interactions between these sectors. Key insights include the need for data standardization and interoperability to enable cross-sector analysis and decision-making. Strengthening utility partnerships is also critical to align energy policies with decarbonization goals and manage the increased demand for renewable energy. The study also emphasizes the importance of equity in the clean energy transition, calling for targeted incentives and support to ensure that low-income and underserved communities benefit from clean technologies like electric vehicles and energy-efficient appliances. To support these efforts, innovative funding mechanisms must be expanded to facilitate interdisciplinary research, such as city-specific decarbonization plans and federal projects like DOE"s Standard Scenarios. By encouraging collaboration and integrating cross-sector insights, this study aims to provide a roadmap to accelerate the clean energy transition and ensure it is both sustainable and inclusive.

14 SOLAR ENERGY↗

Closing the Gap between FAIR Data Repositories and Hierarchical Data Formats

Many in the scientific community, particularly in publicly funded research, are pushing to adhere to more accessible data standards to maximize the findability, accessibility, interoperability, and reusability (FAIR) of scientific data, especially with the growing prevalence of machine learning augmented research. Online FAIR data repositories, such as the Open Science Framework (OSF), help facilitate the adoption of these standards by providing frameworks for storage, access, search, APIs, and other features that create organized hubs of scientific data. However, the wider acceptance of such repositories is hindered by the lack of support of hierarchical data formats, such as Technical Data Management Streaming (TDMS) and Hierarchical Data Format 5 (HDF5), that many researchers rely on to organize their datasets. Various tools and strategies should be used to allow hierarchical data formats, FAIR data repositories, and scientific organizations to work more seamlessly together. A pilot project at Los Alamos National Laboratory (LANL) addresses the disconnect between them by integrating the OSF FAIR data repository with hierarchical data renderers, extending support for additional file types in their framework. The multifaceted interactive renderer displays a tree of metadata alongside a table and plot of the data channels in the file. This allows users to quickly and efficiently load large and complex data files directly in the OSF webapp. Users who are browsing files can quickly and intuitively see the files in the way they or their colleagues structured the hierarchical form and immediately grasp their contents. This solution helps bridge the gap between hierarchical data storage techniques and FAIR data repositories, making both of them more viable options for scientific institutions like LANL which have been put off by the lack of integration between them.

97 MATHEMATICS AND COMPUTING↗

Livewire: Automatic Annotations

Diogenes processes datasets to provide data quality metrics for the Livewire platform and creates standardized data dictionaries from data annotations. Diogenes needs data annotations that clearly outline thenformat and organization of the data. It also relies on the type, class, and unit of each data piece for comprehensive analysis, which it cannot determine independently. The Annotation Tool significantly reduces the time needed to create annotations for Diogenes by generating data annotations with the correct formatting and content. It also employs machine learning and hard-coded models to automatically annotate data class, quality type, and data units.

33 - ADVANCED PROPULSION SYSTEMS↗

VizBrick: A GUI-based Interactive Tool for Authoring Semantic Metadata for Building Datasets

Brick ontology is a unified semantic metadata schema to address the stand-ardization problem of buildings' physical, logical, and virtual assets and the relationships between them. Creating a Brick model for a building dataset means that the dataset's contents are semantically described using the standard terms defined in the Brick ontology. It will enable the benefits of data standardization, without having to recollect or reorganize the data and opens the possibility of automation leveraging the machine readability of the semantic metadata. The problem is that authoring Brick models for building datasets often requires knowledge of semantic technology (e.g., on-tology declarations and RDF syntax) and leads to repeated manual trial and error processes, which can be time-consuming and challenging to do with-out an interactive visual representation of the data. We developed VizBrick, a tool with a graphical user interface that can assist users in creating Brick models visually and interactively without having to understand the Re-source Description Framework (RDF) syntax. VizBrick provides handy ca-pabilities such as keyword search for easy find of relevant brick concepts and relations to their data columns and automatic suggestions of concept mapping. In this demonstration, we present a use-case of VizBrick to show-case how a Brick model can be created for a real-world building dataset.

Lee, Sangkeun (Matt)↗

Measurement of the 252 Cf ⁢(sf) prompt fission neutron spectrum utilizing 12 C ⁡(𝑛, 𝑛) and 9 Be ⁢(𝑛, 𝑛) neutron scattering reference measurements

The 252 Cf spontaneous fission (sf), prompt fission neutron spectrum (PFNS) is a fundamental quantity for nuclear physics measurements of neutron-emitting reactions. This energy distribution of neutrons emitted from fission has been considered a neutron data standard for decades and has been utilized as a reference for neutron detection efficiency, validation of Monte Carlo simulations, benchmarking of dosimetry standards, and more. A significant portion of the global collection of nuclear data on neutron-induced reactions is correlated with the 252 Cf ⁢(sf) PFNS. Despite the reliance on this quantity by the nuclear physics community, the historical collection of 252 Cf PFNS measurements display systematic disagreements that are not understood or easily explained. These experimental discrepancies could potentially bias the 252 Cf PFNS Standard evaluation. On top of this, these past experiments frequently employed correlated experimental measurement or analysis methods. The artificial intelligence (AI)/machine learning (ML)-informed californium chi-nuclear data experiment (AIACHNE) project was formed to (a) investigate these discrepancies utilizing AI/ML methods to identify outlying regions of literature data, assign these regions to features of the experiment itself, and perform an improved evaluation of the 252 Cf PFNS and (b) perform a new experimental measurement of this quantity designed to improve upon the existing literature database. Here, in this work, we report on the AIACHNE 252 Cf PFNS experiment utilizing a new analysis method uncorrelated with all previous measurements: neutron efficiency determinations based on elastic neutron scattering on 12 C and 9 Be . This new method provides an independent test of the existing literature data and evaluation of the 252 Cf ⁢(sf) PFNS. The method is described with detailed covariance quantification procedures, as well as a direct discussion of the sources of uncertainty described as requirements in the “Templates” series of papers. The 252 Cf ⁢(sf) PFNS reported in this work agrees well with the overall shape of the existing standard PFNS evaluation as well as many literature measurements, thus verifying the current evaluation utilizing new techniques. However, the results suggest that there are deficiencies in the angle-differential 12 C and 9 Be ⁢(𝑛, 𝑛) evaluated nuclear data, which produce unphysical structures in the reported result. While these structures are relatively minor, they become obvious because of the high statistical precision of the data and the expected smooth continuity of the 252 Cf ⁢(sf) PFNS.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Towards provision of regularly updated climate data from the Coupled Model Intercomparison Project

The Coupled Model Intercomparison Project (CMIP) is a flagship of the World Climate Research Programme (WCRP). CMIP has become a recognised ‘brand’ in climate circles evolving over the last thirty years from a targeted research activity by a small number of climate modelling centres intercomparing their Earth System Model (ESM) simulations to a broad international coordinated research effort (Durack et al, 2025). CMIP is organized as a research activity leveraging funded and in-kind contributions from experts within modelling centres and the broader scientific community supported more recently by a fully-funded International Project Office. Within CMIP, Model Intercomparison Projects (MIPs) are community-designed to understand past, present and future climate. CMIP data provides a valuable resource for climate research and is routinely used to assess model representation of climate processes and test scientific hypotheses in the context of model uncertainty and (forced and internal) variability as evident from its prolific use in scientific publications1 . The impact relies on enabling infrastructure (most prominently via the Earth System Grid Federation (ESGF)), which allows sharing of simulation output, provision of the boundary conditions used in each simulation, and definition of the data standards that are essential to facilitating wide use of the data. The impact is supplemented by the wide-ranging scrutiny to which model simulations are subjected. Beyond its use in research, CMIP data is a key resource for communities producing derived climate information from downscaling and impact studies, such as the Coordinated Regional Downscaling Experiment (CORDEX; Gutowski et al., 2016) and the Intersectoral Impacts MIP (ISIMIP; Frieler et al., 2024). Government, academic and commercial entities also increasingly rely on CMIP and its downstream data for climate risk assessments and climate services (for example, Copernicus Climate Change Service and World Bank portal). This means that, although CMIP is a research activity, it increasingly serves a secondary and very relevant role as a provider of climate data – a long-recognised dichotomy (Stevens, 2024). Research and applications have distinct needs, with the former requiring flexibility and generality and the latter consistency. Here we explain how the design of the research activity has been adapted to reduce the burdens imposed by applications and how the research infrastructure might evolve to further enable scientific inquiry. We propose one possible approach to consistently providing model information and projections for applications in the future.

Environmental sciences↗

Solar Resource Measurements in Eugene, OR: Cooperative Research and Development Final Report, CRADA Number CRD-07-00252

Site-specific, long-term, continuous, and high-resolution measurements of solar irradiance are important for developing renewable resource data. These data are used for several research and development activities consistent with the NLR mission: establish a national 3-year climatological database of measured solar irradiances; provide high quality ground-truth data for satellite remote sensing validation; support development of radiative transfer models for estimating solar irradiance from available meteorological observations; provide solar resource information needed for technology deployment and operations. Data acquired under this agreement will be available to the public through NLR's Measurement & Instrumentation Data Center – MIDC (http://www.nlr.gov/midc) Or the Renewable Resource Data Center - RReDC (http://rredc.nlr.gov). The MIDC offers a variety of standard data display, access, and analysis tools designed to address the needs of a wide user audience (e.g., industry, academia, and government interests).

14 SOLAR ENERGY↗

radkit v1.2

The radkit software suite (python) consists of three primary libraries: stark, trajan, and curie. The trajan library provides the tools to analyze and manipulate data from lidar and inertial measurement unit (IMU) devices as well as trajectories from algorithms such as simultaneous localization and mapping (SLAM). These components allow reading and writing standard data formats, performing rigid affine transformations, discretizing three-dimensional space, and visualizing data products. The curie library comprises a standard set of object-oriented tools for radiation data and analysis in the following modules: (1) listmode and binmode data classes with methods for manipulation, plotting, slicing and file IO; (2) radiological/nuclear source detection/identification analysis results; (3) source encounters of correlated analyses and (4) energy-dependent angular detector response functions. The stark package provides low-level tools that are leveraged by both curie and trajan. The tools are flexible for offline analysis as well as performant for real-time integrations. The radkit libraries have associated Robot Operating System packages for use in real-time and robotic systems.

Joshi, Tenzing↗

radkit base v1.6

The radkit (base) software suite (python) consists of three primary libraries: stark, trajan, and curie. The trajan library provides the tools to analyze and manipulate data from lidar and inertial measurement unit (IMU) devices, cameras, as well as trajectories from algorithms such as simultaneous localization and mapping (SLAM). These components allow reading and writing standard data formats, performing rigid affine transformations, discretizing three-dimensional space, and visualizing data products. The curie library comprises a standard set of object-oriented tools for radiation data and analysis in the following modules: (1) listmode and binmode data classes with methods for manipulation, plotting, slicing and file IO; (2) radiological/nuclear source detection/identification analysis results; (3) source encounters of correlated analyses and (4) energy-dependent angular detector response functions. The stark package provides low-level tools that are leveraged by both curie and trajan. The tools are flexible for offline analysis as well as performant for real-time integrations.

Salathe, Marco [Lawrence Berkeley National Laborat↗

Legacy Survey of Space and Time Data Preview 2: standard_passband dataset type

We present Rubin Data Preview 2 (DP2), the second data preview from the NDF-DOE Vera C. Rubin Observatory. Data Preview 2 (DP2) comprises coadds, detection catalogs, and ancillary data products; and when fully released will also include single-epoch images and difference images. DP2 is derived from observations acquired by the LSST Science Camera (LSSTCam) on the Simonyi Survey Telescope at the Summit Facility on Cerro Pachón, Chile, primarily during the on-sky commissioning campaign between 2025-04-16 and 2025-09-21, supplemented by observations taken between 2025-10-25 and 2026-01-06 that overlap the commissioning footprint. The DP2 footprint comprises the Science Validation wide-area survey, five Deep Drilling Fields, and a number of targeted small-field regions, including Trifid and Lagoon, Prawn, M49, and New Horizons, all observed as part of the Rubin First Look campaign. Each field was imaged in up to six broad photometric bands, ugrizy, and coadded to produce deep imaging covering an estimated 3,000 deg2. The addition of single-visit-only areas expands the total DP2 footprint to an estimated 15,000 deg2, with coverage in at least one filter. The median per-visit PSF FWHM across the wide-area survey ranges from 1.17 arcsec in the z band to 1.26 arcsec in g and r bands. The deepest field, reaches estimated coadded 5σ depths of u=26 mag, g=26.8 mag, r=26.3 mag, i=26.1 mag, z=25.3 mag, y=23.9 mag. Based on a roughly five-month primary observing baseline and covering only part of the eventual LSST footprint, DP2's area, depth, and multiband coverage nonetheless support a broad range of early science investigations ahead of LSST Data Release This dataset is a subset of the full data release consisting of the standard_passband dataset type. These are the LSSTCam filter bandpasses. This release contains 6 datasets of this type.

79 ASTRONOMY AND ASTROPHYSICS↗

From remotely-sensed solar-induced chlorophyll fluorescence to ecosystem structure, function, and service: Part II—Harnessing data

Although our observing capabilities of solar-induced chlorophyll fluorescence (SIF) have been growing rapidly, the quality and consistency of SIF datasets are still in an active stage of research and development. As a result, there are considerable inconsistencies among diverse SIF datasets at all scales and the widespread applications of them have led to contradictory findings. The present review is the second of the two companion reviews, and data oriented. It aims to (1) synthesize the variety, scale, and uncertainty of existing SIF datasets, (2) synthesize the diverse applications in the sector of ecology, agriculture, hydrology, climate, and socioeconomics, and (3) clarify how such data inconsistency superimposed with the theoretical complexities laid out in may impact process interpretation of various applications and contribute to inconsistent findings. We emphasize that accurate interpretation of the functional relationships between SIF and other ecological indicators is contingent upon complete understanding of SIF data quality and uncertainty. Additionally, biases and uncertainties in SIF observations can significantly confound interpretation of their relationships and how such relationships respond to environmental variations. Built upon our syntheses, we summarize existing gaps and uncertainties in current SIF observations. Further, we offer our perspectives on innovations needed to help improve informing ecosystem structure, function, and service under climate change, including enhancing in-situ SIF observing capability especially in “data desert” regions, improving cross-instrument data standardization and network coordination, and advancing applications by fully harnessing theory and data.

59 BASIC BIOLOGICAL SCIENCES↗