Engineering PapersSearch

SEARCH · Engineering Papers

Results for “Data Base”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 703 records · Page 39

Commissioning of an MPGD-based Drift Chamber for Heavy-ion Tracking in FRIB's Sweeper magnet system

A newly developed drift chamber equipped with an innovative hybrid Micro-Pattern Gaseous Detector based readout was commissioned at FRIB. The detector consists of a Multi-layer Thick Gas Electron Multiplier (M-THGEM) mounted over a high-granularity, position-sensitive readout board. Denoted as the Micro-Pattern Drift Chamber (MPDC), the new device is used to provide tracking capability as part of the detectors of the Sweeper magnet system for neutron-invariant-mass spectrometry at the Facility for Rare Isotope Beams (FRIB). The localization of impinging ions in a 30 × 30 cm 2 drift area is derived by processing the charge-avalanche distribution induced on the segmented readout board. The signals induced on the readout pads are processed by a compact, multi-channel Data Acquisition System (DAQ) based on the Scalable Readout System (SRS). To facilitate synchronization with other detector systems of the Sweeper magnet system, the SRS has been configured to accept an external trigger.

Gaseous detectors

Detailed report on the measurement of the positive muon anomalous magnetic moment to 0.20 ppm

We present details on a new measurement of the muon magnetic anomaly, a μ =(g μ −2)/2. The result is based on positive muon data taken at Fermilab’s Muon Campus during the 2019 and 2020 accelerator runs. The measurement uses 3.1 GeV/c polarized muons stored in a 7.1-m-radius storage ring with a 1.45 T uniform magnetic field. The value of a μ is determined from the measured difference between the muon spin precession frequency and its cyclotron frequency. This difference is normalized to the strength of the magnetic field, measured using nuclear magnetic resonance. The ratio is then corrected for small contributions from beam motion, beam dispersion, and transient magnetic fields. We measure a μ =116592057(25)×10 −11 (0.21 ppm). This is the world’s most precise measurement of this quantity and represents a factor of 2.2 improvement over our previous result based on the 2018 dataset. In combination, the two datasets yield a μ (FNAL)=116592055(24)×10 −11 (0.20 ppm). Combining this with the measurements from Brookhaven National Laboratory for both positive and negative muons, the new world average is a μ (exp)=116592059(22)×10 −11 (0.19 ppm).

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS

Rosenbluth-like separation of the $J/ψ$ near-threshold photoproduction: An access to the gluon gravitational form factors at high t

Here, we perform analysis of the near-threshold $J/\psi $ photoproduction data off the proton based on two theoretical approaches, GPD \cite{Guo3} and holographic \cite{Zahed2}, that represent the differential cross sections as powers of the skewness parameter with coefficients that depend only on the momentum transfer $t$. This allows to separate kinematically the corresponding coefficient functions, in much the same way as this is done for the electric and magnetic form factors using the Rosenbluth separation. We examine the independence of the extracted functions with the photon beam energy. These functions, under additional assumptions, are related to the proton's gluon Gravitational Form Factors (gGFFs). We compare the extracted functions with lattice calculations of the gGFFs in the region of $0.5<|t|<2$~GeV$^{2}$, where they overlap. Such analysis demonstrates the possibility of extracting some combinations of the gGFFs from the data at high $t$, complementary to the lattice calculations available in the low $t$ region. However, higher statistics are needed to more accurately check the predicted scaling behavior of the data and compare with the lattice results, thus testing and comparing the theoretical assumptions used in the GPD and holographic models.

Pentchev, Lubomir [Thomas Jefferson National Accel

Studying baryon acoustic oscillations using photometric redshifts from the DESI Legacy Imaging survey DR9

Context. The Dark Energy Spectroscopic Instrument (DESI) Legacy Imaging Survey DR9 (DR9 hereafter), with its extensive dataset of galaxy locations and photometric redshifts, presents an opportunity to study baryon acoustic oscillations (BAOs) in the region covered by the ongoing spectroscopic survey with DESI. Aims. We aim to investigate differences between different parts of the DR9 footprint. Furthermore, we want to measure the BAO scale for luminous red galaxies within them. Our selected redshift range of 0.6–0.8 corresponds to the bin in which a tension between DESI Y1 and eBOSS was found. Methods. We calculated the anisotropic two-point correlation function in a modified binning scheme to detect the BAOs in DR9 data. We then used template fits based on simulations to measure the BAO scale in the imaging data. Results. Our analysis reveals the expected correlation function shape in most of the footprint areas, showing a BAO scale consistent with Planck’s observations. Aside from identified mask-related data issues in the southern region of the South Galactic Cap, we find a notable variance between the different footprints. Conclusions. We find that this variance is consistent with the difference between the DESI Y1 and eBOSS data, and it supports the argument that that tension is caused by sample variance. Additionally, we also uncovered systematic biases not previously accounted for in photometric BAO studies. We emphasize the necessity of adjusting for the systematic shift in the BAO scale associated with typical photometric redshift uncertainties to ensure accurate measurements.

79 ASTRONOMY AND ASTROPHYSICS

System Engineers and Decisions: It?s All about Knowledge

In order to guarantee that a system meets adequate levels of reliability and availability, system performances are continuously monitored and analyzed thanks to the technological advancements driving the Industry 4.0 revolution. An Industry 4.0 approach is typically based on advanced statistical, big data mining, machine learning, and internet-of-things methods designed to detect anomalies in the behavior of system, detect the most likely failure modes, and provide indications to system engineers on when maintenance activities should be performed before system performance are deemed unacceptable (which can be generated by diagnostic and prognostic methods). However, these analyses, which are designed to automatize and increase the efficacy of the system maintenance program, require large amount of data which can come in various forms: numeric, textual, images, sounds etc. Such data constitutes the historic knowledge benchmark to track system performances and support system engineer decisions. Here we claim that data is not sufficient to support this kind of analyses when applied to systems characterized by complex architectures and behaviors. Robust system engineer decisions require the ability to understand the system operational context that lies behind the observed data elements. In this respect, system models are in fact necessary to “put data in context” and capture relationships between data elements. Industry 4.0 methods require in fact contextual knowledge as a basis upon which hypotheses can be generated and assumptions tested. In our view, for complex systems, model-based system engineering (MBSE) models can afford this contextual knowledge, as they are typically used to describe systems architecture and dynamic behaviors. System knowledge is here intended as the blending of collected data and system architecture which takes the form of a “knowledge graph”. A knowledge graph is a database which consists of a large set of nodes (in our case an entity can be either a data or an MBSE element) which are linked to each other. The types of nodes and links follow a pre-defined topology, sometimes also refers as an ontology, that is designed to fit the actual decisions that needs to be performed. We show here how a knowledge graph can be defined to support system engineer maintenance decisions and how the same graph can be built based on system MBSE models and pre-processed data from numeric (through anomaly detections and diagnostic methods) and textual elements (through technical language processing TLP).

97 - MATHEMATICS AND COMPUTING

DESI DR2 results. II. Measurements of baryon acoustic oscillations and cosmological constraints

We present baryon acoustic oscillation (BAO) measurements from more than 14 million galaxies and quasars drawn from the Dark Energy Spectroscopic Instrument (DESI) Data Release 2 (DR2), based on three years of operation. For cosmology inference, these galaxy measurements are combined with DESI Lyman-𝛼 forest BAO results presented in a companion paper (M. Abdul-Karim et al., companion paper, Phys. Rev. D 112, 083514 2025.). The DR2 BAO results are consistent with DESI DR1 and the Sloan Digital Sky Survey, and their distance-redshift relationship matches those from recent compilations of supernovae (SNe) over the same redshift range. The results are well described by a flat Λ cold dark matter (Λ⁢CDM) model, but the parameters preferred by BAO are in mild, 2.3⁢𝜎 tension with those determined from the cosmic microwave background (CMB), although the DESI results are consistent with the acoustic angular scale 𝜃 * that is well measured by Planck. This tension is alleviated by dark energy with a time-evolving equation of state parametrized by 𝑤0 and 𝑤𝑎, which provides a better fit to the data, with a favored solution in the quadrant with 𝑤 0 >−1 and 𝑤 𝑎 <0. This solution is preferred over Λ ⁢CDM at 3.1⁢𝜎 for the combination of DESI BAO and CMB data. When also including SNe, the preference for a dynamical dark energy model over Λ⁢ CDM ranges from 2.8 − 4.2⁢𝜎 depending on which SNe sample is used. We present evidence from other data combinations which also favor the same behavior at high significance. From the combination of DESI and CMB we derive 95% upper limits on the sum of neutrino masses, finding ∑𝑚 𝜈 < 0.064 eV assuming Λ ⁢CDM and ∑𝑚 𝜈 < 0.16 eV in the 𝑤0⁢𝑤𝑎 model. Unless there is an unknown systematic error associated with one or more datasets, it is clear that Λ⁢ CDM is being challenged by the combination of DESI BAO with other measurements and that dynamical dark energy offers a possible solution.

Baryon acoustic oscillations

TEMPEST

This repository solves the problem of driver identification through vehicular and biometric data. Through an embedding-based approach and a novel loss function, we're able to distinguish between different drivers' behaviors. This also provides preprocessing for reproducibility of results.The code preprocesses vehicular data, trains neural networks, and outputs predictions.This code introduces a novel embedding-based neural network with a 91% rank-1 accuracy, as well as all code to reproduce training and results.

Musgrove, Kyle

Recent Experience with the CMS Data Management System

The CMS[1] experiment manages a large-scale data infrastructure, currently handling over 200 PB of disk and 500 PB of tape storage and transferring more than 1 PB of data per day on average between various WLCG[2] sites. Utilizing Rucio[3] for high-level data management, FTS[4] for data transfers, and a variety of storage and network technologies at the sites, CMS confronts inevitable challenges due to the system’s growing scale and evolving nature. Key challenges include managing transfer and storage failures, optimizing data distribution across different storages based on production and analysis needs, implementing necessary technology upgrades and migrations, and efficiently handling user requests. The data management team has established comprehensive monitoring to supervise this system and has successfully addressed many of these challenges. The team’s efforts aim to ensure data availability and protection, minimize failures and manual interventions, maximize transfer throughput and resource utilization, and provide reliable user support. This paper details the operational experience of CMS with its data management system in recent years, focusing on the encountered challenges, the effective strategies employed to overcome them and the ongoing challenges as we prepare for future demands.

Öztürk, Hasan [CERN]

Autonomous Anomaly Detection For Continuous Streams

The code implements the Isolation Forest (IFML) algorithm within the digital twin (DT) of the AGN-201 nuclear reactor. The DT captures real-time operational data including control rod positions, reactor power, and temperature. The IFML model isolates anomalies by detecting patterns that deviate from expected operational behavior. The algorithm recursively partitions the data and assigns anomaly scores based on the isolation of rare and different events. By tuning parameters specific to the reactor’s operational data, the IFML identifies deviations such as unauthorized material insertions or reactor reactivity shifts. The system streams data using LabView and integrates with the DeepLynx data warehouse for anomaly processing.

Trevino, Eduardo

Building a FAIR data ecosystem for incorporating single-cell transcriptomics data into agricultural genome to phenome research

Introduction The agriculture genomics community has numerous data submission standards available, but the standards for describing and storing single-cell (SC, e.g., scRNA- seq) data are comparatively underdeveloped. Methods To bridge this gap, we leveraged recent advancements in human genomics infrastructure, such as the integration of the Human Cell Atlas Data Portal with Terra, a secure, scalable, open-source platform for biomedical researchers to access data, run analysis tools, and collaborate. In parallel, the Single Cell Expression Atlas at EMBL-EBI offers a comprehensive data ingestion portal for high-throughput sequencing datasets, including plants, protists, and animals (including humans). Developing data tools connecting these resources would offer significant advantages to the agricultural genomics community. The FAANG data portal at EMBL-EBI emphasizes delivering rich metadata and highly accurate and reliable annotation of farmed animals but is not computationally linked to either of these resources. Results Herein, we describe a pilot-scale project that determines whether the current FAANG metadata standards for livestock can be used to ingest scRNA-seq datasets into Terra in a manner consistent with HCA Data Portal standards. Importantly, rich scRNA-seq metadata can now be brokered through the FAANG data portal using a semi-automated process, thereby avoiding the need for substantial expert curation. We have further extended the functionality of this tool so that validated and ingested SC files within the HCA Data Portal are transferred to Terra for further analysis. In addition, we verified data ingestion into Terra, hosted on Azure, and demonstrated the use of a workflow to analyze the first ingested porcine scRNA-seq dataset. Additionally, we have also developed prototype tools to visualize the output of scRNA-seq analyses on genome browsers to compare gene expression patterns across tissues and cell populations. This JBrowse tool now features distinct tracks, showcasing PBMC scRNA-seq alongside two bulk RNA-seq experiments. Discussion We intend to further build upon these existing tools to construct a scientist-friendly data resource and analytical ecosystem based on Findable, Accessible, Interoperable, and Reusable (FAIR) SC principles to facilitate SC-level genomic analysis through data ingestion, storage, retrieval, re-use, visualization, and comparative annotation across agricultural species.

Genetics & Heredity

Assessing Suitable Geologic Carbon Storage Sites Across Utah

Utah has a wealth of potential geological reservoirs for carbon dioxide storage (CS) and a long history of geologic research resulting in an abundance of available subsurface data to evaluate CS potential. Reservoirs may include sandstone, carbonate, and basalt; these rock types are plentiful in Utah’s subsurface and the complex Phanerozoic history throughout the state requires evaluating each geologic region individually for promising reservoir-seal pairs for CO2 storage. Classifying Utah by geologic provinces (or “geo-regions”) allows for customized thinking about suitable CS reservoir and seal distribution, CO2 point sources, land use, and existing infrastructure. Preliminary results from this study highlight the geologic CS potential across 15 geo-regions. Four regions stand out as having high CS potential: the Uinta Basin, San Rafael Swell, Paradox Basin, and the southern Basin and Range Province. The Uinta Basin and San Rafael Swell geo-regions are well suited for CS and have several projects ongoing to evaluate Cretaceous Frontier and Naturita Formations, Jurassic Navajo Sandstone and Entrada Sandstone, and Permian Weber Sandstone reservoir units that lie beneath robust sealing units like the ~5000-ft-thick Mancos Shale and Carmel Formation. Reservoirs such as the Navajo and Weber Sandstones have been demonstrated to be suitable reservoirs through a long history of oil and gas exploration in Utah. New areas of interest include the southern Basin and Range in southwest Utah, where the Jurassic Navajo Sandstone is overlain by the sealing Carmel Formation at suitable depths (>3000 ft), and has good porosities based on outcrop analogue data. Just to the north (e.g., central Basin and Range), legacy wells and 2D seismic data show possible salt and subsurface basalt flows that may provide additional possible CS reservoirs and seals. This geo-region also has the advantage of being coupled with geothermal energy resources that may be used to power burgeoning direct air capture technologies. In the Paradox Basin of southeastern Utah, the Leadville Limestone is a potential storage reservoir beneath the thick (4000–5000 ft), salt-bearing Pennsylvanian Paradox Formation. Although the northern and western parts of Utah offer CS potential, these areas typically contain less infrastructure and subsurface penetrations, creating geologic uncertainty associated with subsurface seals and reservoirs due to a lack of data. Overall, this statewide assessment and ranking is the first step to aid in evaluating CS potential across Utah and provides a foundation for future research in the most favorable locations.

58 GEOSCIENCES

Elucidating Processes Controlling Arctic Atmospheric Aerosol Sources, Aging, and Mixing States (Final Report)

Atmospheric aerosols play critical roles in the Earth’s energy budget, directly by scattering or absorbing solar and terrestrial radiation and indirectly by serving as seeds (nuclei) for cloud droplet and ice crystal formation and by depositing on snow and ice surface, thereby changing the surface albedo. These effects are dependent on aerosol particle size and chemical composition and impact the hydrological cycle as well. This project provided single-particle size and chemical composition measurements across the entire annual cycle in the high Arctic and in the Alaskan Arctic during fall – winter, addressing the most significant gaps in Arctic aerosol observational data. These needs were based on recent rapid sea ice loss across the entire Arctic, as well as the major annual delays in sea ice freeze-up during fall in the Chukchi Sea and increased wintertime sea ice fracturing in the Beaufort Sea, both off the North Slope of Alaska. Two DOE Atmospheric Radiation Measurement (ARM) field campaigns were conducted for atmospheric aerosol sampling. The Aerosols during the Polar Utqiagvik Night (APUN – ‘snow on ground’ in Iñupiaq) ARM field campaign at Utqagivik, Alaska was conducted from Oct. 28 – Dec. 22, 2018. Aerosol sizing instrumentation and a single-particle mass spectrometer were successfully deployed for size-resolved number concentration measurements and measurements of individual particle size and chemical composition, respectively. These results show the influence of locally-produced sea spray aerosol, with high cloud-forming potential, due to delayed sea ice freeze-up in the fall. During the 2019‐2020 international Multidisciplinary drifting Observatory for the Study of Arctic Climate (MOSAiC) expedition, daily atmospheric aerosol particles were collected aboard the German icebreaker Polarstern in the Central Arctic from Nov. 2019 – Oct. 2020. Sea salt aerosol and marine organics were observed year-round during MOSAiC with varying morphologies and sources. These findings are important because most Arctic models do not include a sea spray aerosol source, despite this source increasing with declining sea ice extent. In addition to collecting new samples and data, this project also conducted further analysis of previously collected single-particle chemical composition measurements within the North Slope of Alaska oil fields and at Utqiaġvik, AK, during Aug. – Sep. 2015 and 2016 field campaigns. This work resulted in the discovery of chemical reactions of oil field combustion emissions occurring within fog droplets across the North Slope of Alaska and forming secondary aerosol, showing the impact of Arctic oil field emissions beyond black carbon aerosol and greenhouse gases. In addition, the distribution of chemical species across the aerosol population within the oil fields was quantified, using these data and a previously development framework. We also presented the first ambient evidence of the collision of two atmospheric particles resulting in formation of an organic-coated ammonium sulfate particle of marine origin, which has implications for cloud formation with declining sea ice extent. Overall, this project has elucidated connections between seawater biogeochemistry, resource extraction activities, atmospheric composition, clouds, and the energy budget of the Arctic region. The results of this project are expected to improve weather and sea ice forecasting for security and development in the Arctic and beyond.

54 ENVIRONMENTAL SCIENCES

Label-based Virtual Directories In dCache

Traditional filesystems organize data in directories. These directories are typically a collection of files whose grouping is based on a single criterion, e.g., the starting date of an experiment, experiment name, beamline ID, measurement device, or instrument. However, each file in a directory can belong to several logical groups, such as a special event type, experiment condition, or a part of a selected dataset. dCache is a storage system developed to store large amounts of scientific data, used by many HEP and Photon Science experiments. With recent developments in dCache, we have introduced a concept of file tagging, which dynamically groups files with the same label into virtual directories. The file labels can be added, removed, renamed, and deleted through the admin interface or via REST API. The files in virtual directories are exposed through all protocols supported by dCache. This contribution will describe the details of the implementation for file tagging in dCache and present our future development plans on automatic metadata extractions, a feature that will significantly simplify data management. Additionally, we are exploring the future use of virtual directories as a way to translate scientific data catalogs into filesystem views for direct data analysis.

Sahakyan, Marina [DESY]

Neutron powder diffraction, Mossbauer Spectroscopy and Optical Spectroscopy to study magnetic and nuclear lattices in Fe-based oxychlorides Ca2FeO3Cl, Sr2FeO3Cl and Sr3Fe2O5Cl2

Data for Neutron Diffraction, Mossbauer Spectoscopy and Optical Spectroscopy are contained on all three samples (Ca2FeO3Cl, Sr2FeO3Cl, and Sr3Fe2O5Cl2). Mossbauer data are in the folder "MossbauerSpectroscopy". This contains details of files and examples to read the data. The Optical Spectroscopy are in the folder "AbsorptionData", this contains one file with explanatory headers. The neutron diffraction data are in the "NeutronDiffractionData" folder. The data was collected on the HB-2A powder diffractometer at HFIR. The autoreduced files are corrected using a vanadium standard and are in arbitrary intensity units. The RawData folder contains the uncorrected data with metadata for motor positions and temperature. All measurements were collected with a constant neuton wavelength of 2.41 Angstrom. An excel spreadsheet "NeutronDiffractionData_IPTS-29118_summary" contains details for each scan. Sr3Fe2O5Cl2 data were collected at only room temperature. Ca2FeO3Cl and Sr2FeO3Cl data were collected at 4K and room temperature. The autoreduced .dat fileformat is three columns corresponding to: Two-theta, Intensity and Intensity_error.

fe-based oxychlorides

Mountain Basin Controls on the Snow-to-Streamflow Signal: An AIC-Weighted Multiple Linear Regression Framework

A regression-based analysis quantifies how basin characteristics modulate the snow-to-streamflow signal. First, we use the ERA5-Land reanalysis gridded product (European Centre for Medium Range Weather Forecasts reanalysis 5 -Land component) for 4,655 hydrologic unit code - 10 (HUC10) mountain basins across the western United States (US) for water years 1987–2024. Linear regressions are performed for peak snow water equivalent (SWE) and annual streamflow for each mountain basin. Models use ordinary least squares in Python’s statsmodels package. After which, an Akaike Information Criterion (AIC)–weighted ensemble multiple linear regression (MLR) framework with 47 watershed traits is used to predict the linear regression coefficient of determination (r-squared) defining the ability of peak SWE to predict annual streamflow across all mountain basin. Predictor sets are constrained to avoid multicollinearity by excluding models with variance inflation factors (VIF) greater than 5. Mountain basin traits included in the MLR include seasonal climate, topography, vegetation type and structure, and bedrock geology. Accepted models are considered if their AIC is within 2.0 of the model with the minimum AIC, or best model. To compare predictor influence across acceptable models, we computed standardized regression coefficients. To evaluate structural redundancy among models, we constructed binary inclusion vectors for each acceptable model, denoting whether a predictor was present (1) or absent (0). Core predictor variables are defined as occurring in at least 67% of the acceptable models. For this regional analysis, only one model was found acceptable, with higher snow-to-streamflow translation (higher r-squared) occurring in colder mountain basins with higher relative winter precipitation, more snow accumulation and a lower fraction of annual precipitation that falls in the spring and summer. The second component of the data package uses previously published, high-resolution output from an integrated hydrological model of the East River watershed using the U.S. Geological Survey Groundwater and Surface water Flow model (GSFLOW, doi:10.15485/1998576). East River MLR expands upon the approach described above to explore the response of five streamflow metrics—annual streamflow, runoff efficiency, 7-day minimum flow, low-flow duration, and non-perennial stream fraction to snow system indicators including peak SWE, snow-covered area, snow disappearance date, and the fraction of basin area characterized by low-to-no snow, as well as seasonal precipitation and temperature, and annual hydrologic variables representing soil moisture, evapotranspiration (ET), the partitioning of incoming precipitation to evapotranspiration (ET/P), groundwater storage, and groundwater inflow to streams. MLR was done on all water years (P0: 1987-2024) and for each period as determined in the split analysis using pooled regression techniques (P1: 1987-2011 and P2: 2012-2024) to evaluate shifting predictor variable emphasis on streamflow generation. Results indicate that since 2012, peak SWE has lost statistical strength in its prediction of annual streamflow and runoff efficiency, and the indirect influence of spring temperature has emerged as critically important. Low-flow metrics remain largely influenced by soil moisture, vegetation water use and groundwater inflows with summer precipitation becoming a direct influence on minimum summer flow. Together, these data and Python-based analysis tools provide a framework for identifying the key watershed characteristics that control how streamflow responds to snow from year to year. The package also helps quantify uncertainty in statistical models and assess how snow–streamflow relationships vary across regions and over time. This dataset contains comma-separated values files (.csv), text files (.txt), python code files (.py), figure files (.png), and shapefiles (.cpg, .dbf, .prj, .sbn, .sbx, .shp, .xml). Further details on file contents and MLR execution can be found in the readme file and the FLMD files. Work was supported by the Watershed Function Science Focus Area at Lawrence Berkeley National Laboratory funded by the US Department of Energy, Office of Science, Biological and Environmental Research under Contract No. DE-AC02-05CH11231.

54 ENVIRONMENTAL SCIENCES

SlimIO: Lightweight I/O Path Design for Write Isolation in FDP-backed In-Memory Databases

In-Memory Databases (IMDBs) are widely used with HPC applications to manage transient data, often using snapshot-based persistence for backups. Redis, a representative IMDB, employs both snapshot and Write-Ahead Log (WAL) mechanisms, storing data on persistent devices via the traditional kernel I/O path. This method incurs syscall overhead, I/O contention between processes, and SSD garbage collection (GC) delays. To address these issues, we propose SlimIO, which adopts I/O passthru to minimize syscall overhead and inter-process I/O interference. Additionally, it leverages Flexible Data Placement (FDP) SSDs as backup storage to avoid performance degradation from SSD GC. Experimental results show that SlimIO reduces snapshot time by up to 25%, increases query throughput by up to 30% during non-snapshot periods, and lowers 99.9%-ile latency by up to 50%. Furthermore, it achieves a write amplification factor (WAF) of 1.00, indicating no redundant internal writes, thus extending SSD lifespan.

Lee, Sangyun [Sogang University]

Synoptic Weather Regime Classifications for June, July, August, and September, 2022

The synoptic weather regime classification has become a highly demanded product for the ARM site in recent years. This type of regime classification has shown applications in various studies and topics, including aerosol-cloud interactions, land-atmosphere interactions, and cloud radiative effects. The VAP employs an unsupervised machine learning method, Self-organizing map (SOM), to classify weather regimes for each day of the AMF campaigns and fixed sites, using ERA5 data. This idea is mainly based on our published study for TRACER in Wang et al. (2022, JGR-A). This dataset includes the data in June, July, August, and September; the last year of the data is 2022.

54 ENVIRONMENTAL SCIENCES

DeepLynx Ecosystem 2025

Poor data integration and governance continue to plague complex engineering projects, resulting in missed cost, schedule, and performance targets. Departments operate in isolated systems with manual data exchange, creating fragmented information that compounds errors and leads to significant delays and cost overruns. The DeepLynx ecosystem addresses these challenges through an open-source, modular data management platform that transforms fragmented project data into an integrated digital thread. Built on a federated microservice architecture, the ecosystem comprises seven specialized tools centered around DeepLynx Nexus, a unified data catalog with hierarchical organization and graph-based navigation capabilities. The ecosystem includes: DeepLynx Stream for real-time timeseries data ingestion from industrial sources; DeepLynx Ingest for governed data uploads with formal review workflows; DeepLynx Lattice for ontology-based entity and relationship extraction; DeepLynx Run for workflow orchestration and secure AI/ML compute; DeepLynx Visualize for 3D digital twin visualization; and DeepLynx Insight for AI-assisted document analysis with traceable, grounded responses. Deployable in cloud, on-premise, or hybrid environments using containerized Docker applications and Helm charts, the DeepLynx ecosystem provides flexible infrastructure that adapts to organizational requirements. By consolidating project data into a unified data lake with role-based access controls and OAuth2 authentication, DeepLynx enables digital thread and digital twin capabilities that improve decision-making, reduce risk, and support complex engineering workflows throughout the project lifecycle.

42 - ENGINEERING