Engineering PapersSearch

SEARCH · Engineering Papers

Results for “Data Base”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 649 records · Page 36

Legacy Survey of Space and Time Data Preview 1: ss_object dataset type

The Legacy Survey of Space and Time Data Preview 1 (DP1) is the first release of data from the NSF-DOE Vera C. Rubin Observatory. It consists of raw and calibrated single-epoch images, co-adds, difference images, detection catalogs, and other derived data products. DP1 is based on 1792 science-grade optical/near-infrared exposures acquired over 48 distinct nights by the Rubin Commissioning Camera, LSSTComCam, on the Simonyi Survey Telescope at the Summit Facility on Cerro Pachón, Chile during the first on-sky commissioning campaign in late 2024. DP1 covers a total of approximately 15 sq. deg. over seven roughly equally-sized non-contiguous fields, each independently observed in six broad photometric bands, ugrizy, spanning a range of stellar densities and latitudes and overlapping with external reference datasets. This dataset is a subset of the full data release consisting of the ss_object dataset type. These are derived parameters for moving objects. This release contains 1 dataset of this type.

79 ASTRONOMY AND ASTROPHYSICS

Legacy Survey of Space and Time Data Preview 1: SSObject searchable catalog

The Legacy Survey of Space and Time Data Preview 1 (DP1) is the first release of data from the NSF-DOE Vera C. Rubin Observatory. It consists of raw and calibrated single-epoch images, co-adds, difference images, detection catalogs, and other derived data products. DP1 is based on 1792 science-grade optical/near-infrared exposures acquired over 48 distinct nights by the Rubin Commissioning Camera, LSSTComCam, on the Simonyi Survey Telescope at the Summit Facility on Cerro Pachón, Chile during the first on-sky commissioning campaign in late 2024. DP1 covers a total of approximately 15 sq. deg. over seven roughly equally-sized non-contiguous fields, each independently observed in six broad photometric bands, ugrizy, spanning a range of stellar densities and latitudes and overlapping with external reference datasets. This dataset is a subset of the full data release consisting of a searchable catalog named SSObject. This catalog contains derived parameters for moving objects. This catalog contains 431 rows with 3 columns.

79 ASTRONOMY AND ASTROPHYSICS

Legacy Survey of Space and Time Data Preview 1: standard_bandpass dataset type

The Legacy Survey of Space and Time Data Preview 1 (DP1) is the first release of data from the NSF-DOE Vera C. Rubin Observatory. It consists of raw and calibrated single-epoch images, co-adds, difference images, detection catalogs, and other derived data products. DP1 is based on 1792 science-grade optical/near-infrared exposures acquired over 48 distinct nights by the Rubin Commissioning Camera, LSSTComCam, on the Simonyi Survey Telescope at the Summit Facility on Cerro Pachón, Chile during the first on-sky commissioning campaign in late 2024. DP1 covers a total of approximately 15 sq. deg. over seven roughly equally-sized non-contiguous fields, each independently observed in six broad photometric bands, ugrizy, spanning a range of stellar densities and latitudes and overlapping with external reference datasets. This dataset is a subset of the full data release consisting of the standard_bandpass dataset type. These are the LSSTComCam filter bandpasses. This release contains 6 datasets of this type.

79 ASTRONOMY AND ASTROPHYSICS

Evaluation of Physical Microphysical Property Retrieval Algorithms During the 2020 IMPACTS Field Campaign

The NASA Investigation of Microphysics and Precipitation for Atlantic Coast Threatening Snowstorms (IMPACTS) field campaign provides high-quality, high-altitude aircraft lidar (532 nm), radar (W-band) and in-cloud microphysical aircraft data taken during wintertime storm events impacting the United States. This study evaluates two mass-dimensional relationships (Brown and Francis (1995, BF95); Heymsfield (2014, H14) and two lidar-radar microphysical retrieval algorithms (Cloudsat and CALIPSO Ice Cloud Property Product (2C-ICE); VarPy (a variational method derived from the satellite lidar-radar data community)) to estimate aircraft-retrieved volume extinction coefficient (σ), ice water content (IWC), and effective radius (r e ) during the 2020 IMPACTS deployment. BF95 and H14 have a close 1:1 correlation (R 2 = 0.98) with in-situ observations of σ. However, only BF95 displays a linear, consistent, and almost temperature-independent low bias for IWC and r e , which likely arises from the environmental conditions used to determine each. Unlike the field-campaign-derived BF95 and H14 relationships, VarPy and 2C-ICE directly ingest the aircraft-based lidar and radar data to simulate σ, IWC, and r e . For all three microphysical parameters, VarPy and 2C-ICE retrieval errors became notably more pronounced around the dendritic growth zone (-15°C to -10°C) and near freezing (≥-5°C), which suggests that both algorithms experience difficulty addressing riming and aggregation processes and with larger particles (dendrites and plates) due in part to their simplified ice particle assumptions. However, the mean-melt diameter ice-particle assumption did yield more accurate IWC estimates, which led to slightly better overall results for VarPy.

54 ENVIRONMENTAL SCIENCES

Using Temporal Information from Human Mobility Data to Detect Anchor Points

Spatiotemporal mobility data are available in massive quantities, but large quantities of data typically include fewer variables or data fields. Often, the only available fields are User ID, Longitude, Latitude, Timestamp (ULLT). This raises an important question: how much can we infer about human mobility patterns using only these four fields? With ULLT data, we do not know individuals' socioeconomic status information or when they are visiting their anchor points (AP) or locations (such as homes, places of employment, or schools), and it is a modern challenge to use this data to infer these characteristics. When detecting anchor locations with limited input information, verification and validation (VV) are significant challenges. This paper addresses the problem of identifying individuals' anchor locations using only temporal information from spatiotemporal datasets with limited attributes. Our approach does not explicitly use latitude and longitude during analysis. Locationbased information is only employed in the preprocessing stage to identify periods of movement (trips) and stops (dwelling). Beyond this step, all analysis is based on temporal patterns. In theory, if stops and dwell times could be detected through alternative means, our method could function entirely without location-based input. We demonstrate this methodology on the 2017 National Household Travel Survey (NHTS) data, because it includes a carefully designed and collected time use survey with representative sampling and labeled ground truth. The high-quality survey data allows us to test the accuracy of our methods because NHTS contains intended place labels and agent/user characteristics. We have also applied our validated AP identification algorithm on very large-scale GPS based trajectory data for Patterns-of-Life (PoL) assessment and other applications, but due to space limit that could not be presented here.

McBride, Liz [ORNL] (ORCID:0000000286925869)

CO2-Locate: A Dynamic Database and Tool for Accessing National Oil and Gas Well Data to Inform Carbon Storage Projects

The CO2-Locate Database is a growing compilation of publicly available wellbore resources that have been merged based on common attributes across data sources with an attribute schema developed to be consistent across disparate resources, reduce data gaps, and eliminate record redundancy. The first version of CO2-Locate has been published to Energy Data eXchange (EDX) and includes the integrated public wells dataset as well as additional geospatial summary layers of key wellbore characteristics to protect proprietary resources. Additionally, the CO2-Locate database has been deployed into a web application, enabling easy access, data filtering capabilities, and visualization of U.S. wellbore infrastructure by stakeholders to inform injection site selection and risk assessments.

Dyer, Alec S. [NETL Site Support Contractor, Natio

Near-Efficient and Non-Asymptotic Multiway Inference

We establish non-asymptotic efficiency guarantees for tensor decomposition–based inference in count data models. Under a Poisson framework, we consider two related goals: (i) parametric inference , the estimation of the full distributional parameter tensor, and (ii) multiway analysis , the recovery of its canonical polyadic (CP) decomposition factors. Our main result shows that in the rank-one setting, a rank-constrained maximum-likelihood estimator achieves multiway analysis with variance matching the Cramér–Rao Lower Bound (CRLB) up to absolute constants and logarithmic factors. This provides a general framework for studying “near-efficient” multiway estimators in finite-sample settings. For higher ranks, we illustrate that our multiway estimator may not attain the CRLB; nevertheless, CP-based parametric inference remains nearly minimax optimal, with error bounds that improve on prior work by offering more favorable dependence on the CP rank. Numerical experiments corroborate near-efficiency in the rank-one case and highlight the efficiency gap in higher-rank scenarios.

97 MATHEMATICS AND COMPUTING

Availability of Critical Benchmark Experiments for the Pebble Tanker Transportation Model for Nuclear Criticality Safety Validation of TRISO Pebbles

This study addresses the need for comprehensive investigations into TRi-structural ISOtropic (TRISO) fuel pebble transportation validation. In this work, an exploratory model, the pebble tanker(PT), was developed with the aim of facilitating the validation of nuclear criticality safety calculations in the context of industrial-scale transportation of TRISO fuel. The PT model was designed to investigate the availability and applicability of critical benchmark experiments crucial for assessing the transportation of these pebbles. This work incorporated sensitivity/uncertainty (S/U) similarity studies to quantify the applicability of critical benchmark experiments and to address nuclear data uncertainties in the context of TRISO transportation. Two container models were investigated: one for the Hermes-type pebble and one for the Pebble Bed Modular Reactor (PBMR)–type pebble. The models were simplified, considering fuel, containment, and either water or air, to enable a focus on the underlying physics of applications involving TRISO fuel pebbles using the PT model. A crucial aspect under consideration was the capacity of the transport package to hold pebbles while ensuring subcriticality in the flooded state. An approach in the criticality validation process involves assessing the similarity between systems through an integral index parameter evaluation. This involves calculating a correlation coefficient (referred to as c k ) based on shared nuclear data–induced uncertainty between a benchmark experiment and the application of the PT model. To facilitate this analysis, the SCALE tools, particularly the CSAS6-Shift, TSUNAMI-3D-Shift, and TSUNAMI-IP sequences, were employed for comprehensive studies in neutronics and S/U analysis. Our findings showed that there are sufficient critical experimental benchmarks to perform this validation of the PT model in the most reactive state, i.e. when the tanker is flooded. This paper provides valuable insights into validating a transport package for Generation IV TRISO fuel pebbles.

22 GENERAL STUDIES OF NUCLEAR REACTORS

Commissioning of an MPGD-based Drift Chamber for Heavy-ion Tracking in FRIB's Sweeper magnet system

A newly developed drift chamber equipped with an innovative hybrid Micro-Pattern Gaseous Detector based readout was commissioned at FRIB. The detector consists of a Multi-layer Thick Gas Electron Multiplier (M-THGEM) mounted over a high-granularity, position-sensitive readout board. Denoted as the Micro-Pattern Drift Chamber (MPDC), the new device is used to provide tracking capability as part of the detectors of the Sweeper magnet system for neutron-invariant-mass spectrometry at the Facility for Rare Isotope Beams (FRIB). The localization of impinging ions in a 30 × 30 cm 2 drift area is derived by processing the charge-avalanche distribution induced on the segmented readout board. The signals induced on the readout pads are processed by a compact, multi-channel Data Acquisition System (DAQ) based on the Scalable Readout System (SRS). To facilitate synchronization with other detector systems of the Sweeper magnet system, the SRS has been configured to accept an external trigger.

Gaseous detectors

Detailed report on the measurement of the positive muon anomalous magnetic moment to 0.20 ppm

We present details on a new measurement of the muon magnetic anomaly, a μ =(g μ −2)/2. The result is based on positive muon data taken at Fermilab’s Muon Campus during the 2019 and 2020 accelerator runs. The measurement uses 3.1 GeV/c polarized muons stored in a 7.1-m-radius storage ring with a 1.45 T uniform magnetic field. The value of a μ is determined from the measured difference between the muon spin precession frequency and its cyclotron frequency. This difference is normalized to the strength of the magnetic field, measured using nuclear magnetic resonance. The ratio is then corrected for small contributions from beam motion, beam dispersion, and transient magnetic fields. We measure a μ =116592057(25)×10 −11 (0.21 ppm). This is the world’s most precise measurement of this quantity and represents a factor of 2.2 improvement over our previous result based on the 2018 dataset. In combination, the two datasets yield a μ (FNAL)=116592055(24)×10 −11 (0.20 ppm). Combining this with the measurements from Brookhaven National Laboratory for both positive and negative muons, the new world average is a μ (exp)=116592059(22)×10 −11 (0.19 ppm).

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS

Rosenbluth-like separation of the $J/ψ$ near-threshold photoproduction: An access to the gluon gravitational form factors at high t

Here, we perform analysis of the near-threshold $J/\psi $ photoproduction data off the proton based on two theoretical approaches, GPD \cite{Guo3} and holographic \cite{Zahed2}, that represent the differential cross sections as powers of the skewness parameter with coefficients that depend only on the momentum transfer $t$. This allows to separate kinematically the corresponding coefficient functions, in much the same way as this is done for the electric and magnetic form factors using the Rosenbluth separation. We examine the independence of the extracted functions with the photon beam energy. These functions, under additional assumptions, are related to the proton's gluon Gravitational Form Factors (gGFFs). We compare the extracted functions with lattice calculations of the gGFFs in the region of $0.5<|t|<2$~GeV$^{2}$, where they overlap. Such analysis demonstrates the possibility of extracting some combinations of the gGFFs from the data at high $t$, complementary to the lattice calculations available in the low $t$ region. However, higher statistics are needed to more accurately check the predicted scaling behavior of the data and compare with the lattice results, thus testing and comparing the theoretical assumptions used in the GPD and holographic models.

Pentchev, Lubomir [Thomas Jefferson National Accel

Studying baryon acoustic oscillations using photometric redshifts from the DESI Legacy Imaging survey DR9

Context. The Dark Energy Spectroscopic Instrument (DESI) Legacy Imaging Survey DR9 (DR9 hereafter), with its extensive dataset of galaxy locations and photometric redshifts, presents an opportunity to study baryon acoustic oscillations (BAOs) in the region covered by the ongoing spectroscopic survey with DESI. Aims. We aim to investigate differences between different parts of the DR9 footprint. Furthermore, we want to measure the BAO scale for luminous red galaxies within them. Our selected redshift range of 0.6–0.8 corresponds to the bin in which a tension between DESI Y1 and eBOSS was found. Methods. We calculated the anisotropic two-point correlation function in a modified binning scheme to detect the BAOs in DR9 data. We then used template fits based on simulations to measure the BAO scale in the imaging data. Results. Our analysis reveals the expected correlation function shape in most of the footprint areas, showing a BAO scale consistent with Planck’s observations. Aside from identified mask-related data issues in the southern region of the South Galactic Cap, we find a notable variance between the different footprints. Conclusions. We find that this variance is consistent with the difference between the DESI Y1 and eBOSS data, and it supports the argument that that tension is caused by sample variance. Additionally, we also uncovered systematic biases not previously accounted for in photometric BAO studies. We emphasize the necessity of adjusting for the systematic shift in the BAO scale associated with typical photometric redshift uncertainties to ensure accurate measurements.

79 ASTRONOMY AND ASTROPHYSICS

System Engineers and Decisions: It?s All about Knowledge

In order to guarantee that a system meets adequate levels of reliability and availability, system performances are continuously monitored and analyzed thanks to the technological advancements driving the Industry 4.0 revolution. An Industry 4.0 approach is typically based on advanced statistical, big data mining, machine learning, and internet-of-things methods designed to detect anomalies in the behavior of system, detect the most likely failure modes, and provide indications to system engineers on when maintenance activities should be performed before system performance are deemed unacceptable (which can be generated by diagnostic and prognostic methods). However, these analyses, which are designed to automatize and increase the efficacy of the system maintenance program, require large amount of data which can come in various forms: numeric, textual, images, sounds etc. Such data constitutes the historic knowledge benchmark to track system performances and support system engineer decisions. Here we claim that data is not sufficient to support this kind of analyses when applied to systems characterized by complex architectures and behaviors. Robust system engineer decisions require the ability to understand the system operational context that lies behind the observed data elements. In this respect, system models are in fact necessary to “put data in context” and capture relationships between data elements. Industry 4.0 methods require in fact contextual knowledge as a basis upon which hypotheses can be generated and assumptions tested. In our view, for complex systems, model-based system engineering (MBSE) models can afford this contextual knowledge, as they are typically used to describe systems architecture and dynamic behaviors. System knowledge is here intended as the blending of collected data and system architecture which takes the form of a “knowledge graph”. A knowledge graph is a database which consists of a large set of nodes (in our case an entity can be either a data or an MBSE element) which are linked to each other. The types of nodes and links follow a pre-defined topology, sometimes also refers as an ontology, that is designed to fit the actual decisions that needs to be performed. We show here how a knowledge graph can be defined to support system engineer maintenance decisions and how the same graph can be built based on system MBSE models and pre-processed data from numeric (through anomaly detections and diagnostic methods) and textual elements (through technical language processing TLP).

97 - MATHEMATICS AND COMPUTING

DESI DR2 results. II. Measurements of baryon acoustic oscillations and cosmological constraints

We present baryon acoustic oscillation (BAO) measurements from more than 14 million galaxies and quasars drawn from the Dark Energy Spectroscopic Instrument (DESI) Data Release 2 (DR2), based on three years of operation. For cosmology inference, these galaxy measurements are combined with DESI Lyman-𝛼 forest BAO results presented in a companion paper (M. Abdul-Karim et al., companion paper, Phys. Rev. D 112, 083514 2025.). The DR2 BAO results are consistent with DESI DR1 and the Sloan Digital Sky Survey, and their distance-redshift relationship matches those from recent compilations of supernovae (SNe) over the same redshift range. The results are well described by a flat Λ cold dark matter (Λ⁢CDM) model, but the parameters preferred by BAO are in mild, 2.3⁢𝜎 tension with those determined from the cosmic microwave background (CMB), although the DESI results are consistent with the acoustic angular scale 𝜃 * that is well measured by Planck. This tension is alleviated by dark energy with a time-evolving equation of state parametrized by 𝑤0 and 𝑤𝑎, which provides a better fit to the data, with a favored solution in the quadrant with 𝑤 0 >−1 and 𝑤 𝑎 <0. This solution is preferred over Λ ⁢CDM at 3.1⁢𝜎 for the combination of DESI BAO and CMB data. When also including SNe, the preference for a dynamical dark energy model over Λ⁢ CDM ranges from 2.8 − 4.2⁢𝜎 depending on which SNe sample is used. We present evidence from other data combinations which also favor the same behavior at high significance. From the combination of DESI and CMB we derive 95% upper limits on the sum of neutrino masses, finding ∑𝑚 𝜈 < 0.064 eV assuming Λ ⁢CDM and ∑𝑚 𝜈 < 0.16 eV in the 𝑤0⁢𝑤𝑎 model. Unless there is an unknown systematic error associated with one or more datasets, it is clear that Λ⁢ CDM is being challenged by the combination of DESI BAO with other measurements and that dynamical dark energy offers a possible solution.

Baryon acoustic oscillations

TEMPEST

This repository solves the problem of driver identification through vehicular and biometric data. Through an embedding-based approach and a novel loss function, we're able to distinguish between different drivers' behaviors. This also provides preprocessing for reproducibility of results.The code preprocesses vehicular data, trains neural networks, and outputs predictions.This code introduces a novel embedding-based neural network with a 91% rank-1 accuracy, as well as all code to reproduce training and results.

Musgrove, Kyle

Recent Experience with the CMS Data Management System

The CMS[1] experiment manages a large-scale data infrastructure, currently handling over 200 PB of disk and 500 PB of tape storage and transferring more than 1 PB of data per day on average between various WLCG[2] sites. Utilizing Rucio[3] for high-level data management, FTS[4] for data transfers, and a variety of storage and network technologies at the sites, CMS confronts inevitable challenges due to the system’s growing scale and evolving nature. Key challenges include managing transfer and storage failures, optimizing data distribution across different storages based on production and analysis needs, implementing necessary technology upgrades and migrations, and efficiently handling user requests. The data management team has established comprehensive monitoring to supervise this system and has successfully addressed many of these challenges. The team’s efforts aim to ensure data availability and protection, minimize failures and manual interventions, maximize transfer throughput and resource utilization, and provide reliable user support. This paper details the operational experience of CMS with its data management system in recent years, focusing on the encountered challenges, the effective strategies employed to overcome them and the ongoing challenges as we prepare for future demands.

Öztürk, Hasan [CERN]

Autonomous Anomaly Detection For Continuous Streams

The code implements the Isolation Forest (IFML) algorithm within the digital twin (DT) of the AGN-201 nuclear reactor. The DT captures real-time operational data including control rod positions, reactor power, and temperature. The IFML model isolates anomalies by detecting patterns that deviate from expected operational behavior. The algorithm recursively partitions the data and assigns anomaly scores based on the isolation of rare and different events. By tuning parameters specific to the reactor’s operational data, the IFML identifies deviations such as unauthorized material insertions or reactor reactivity shifts. The system streams data using LabView and integrates with the DeepLynx data warehouse for anomaly processing.

Trevino, Eduardo

Building a FAIR data ecosystem for incorporating single-cell transcriptomics data into agricultural genome to phenome research

Introduction The agriculture genomics community has numerous data submission standards available, but the standards for describing and storing single-cell (SC, e.g., scRNA- seq) data are comparatively underdeveloped. Methods To bridge this gap, we leveraged recent advancements in human genomics infrastructure, such as the integration of the Human Cell Atlas Data Portal with Terra, a secure, scalable, open-source platform for biomedical researchers to access data, run analysis tools, and collaborate. In parallel, the Single Cell Expression Atlas at EMBL-EBI offers a comprehensive data ingestion portal for high-throughput sequencing datasets, including plants, protists, and animals (including humans). Developing data tools connecting these resources would offer significant advantages to the agricultural genomics community. The FAANG data portal at EMBL-EBI emphasizes delivering rich metadata and highly accurate and reliable annotation of farmed animals but is not computationally linked to either of these resources. Results Herein, we describe a pilot-scale project that determines whether the current FAANG metadata standards for livestock can be used to ingest scRNA-seq datasets into Terra in a manner consistent with HCA Data Portal standards. Importantly, rich scRNA-seq metadata can now be brokered through the FAANG data portal using a semi-automated process, thereby avoiding the need for substantial expert curation. We have further extended the functionality of this tool so that validated and ingested SC files within the HCA Data Portal are transferred to Terra for further analysis. In addition, we verified data ingestion into Terra, hosted on Azure, and demonstrated the use of a workflow to analyze the first ingested porcine scRNA-seq dataset. Additionally, we have also developed prototype tools to visualize the output of scRNA-seq analyses on genome browsers to compare gene expression patterns across tissues and cell populations. This JBrowse tool now features distinct tracks, showcasing PBMC scRNA-seq alongside two bulk RNA-seq experiments. Discussion We intend to further build upon these existing tools to construct a scientist-friendly data resource and analytical ecosystem based on Findable, Accessible, Interoperable, and Reusable (FAIR) SC principles to facilitate SC-level genomic analysis through data ingestion, storage, retrieval, re-use, visualization, and comparative annotation across agricultural species.

Genetics & Heredity