Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “open data format”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 127 records · Page 7

Dissolution zone model of the oxide structure in additively manufactured dispersion-strengthened alloys

The structural evolution of oxides in dispersion-strengthened superalloys during laser-powder bed fusion is considered in detail. Alloy chemistry and process parameter effects on oxide structure are assessed through a parameter study on the model alloy Ni-20Cr, doped with varying concentrations of Y 2 O 3 and Al. Small angle neutron scattering measurements of the dispersoid size distribution show the dispersoid size increases with higher laser power, slower scan speed, and increasing Y 2 O 3 and Al content. Complementary electron microscopy measurements reveal reactions between Y 2 O 3 and Al, even in nanoscale dispersoids, and the presence of micron-scale oxide slag inclusions in select specimens. A scaling analysis of mass and momentum transport within the melt pool, presented here, establishes that diffusional structural evolution mechanisms dominate for nanoscale dispersoids, while fluid forces and advection become significant for larger slag inclusions. These findings are developed into a theory of dispersoid structural evolution, integrating quantitative models of diffusional processes – dispersoid dissolution, nucleation, growth, coarsening – with a reduced order model of time-temperature trajectories of fluid parcels within the melt pool. Calculations of the dispersoid size in single-pass melting reveal a zone in the center of the melt track in which the oxide feedstock fully dissolves. Within this zone the final Y 2 O 3 size is independent of feedstock size and determined by nucleation and growth kinetics. If the dissolution zones of adjacent melt tracks overlap sufficiently with each other to dissolve large oxides, formed during printing or present in the powder feedstock, then the dispersoid structure throughout the build volume is homogeneous and matches that from a single pass within the dissolution zone. Gaps between adjacent dissolution zones result in oxide accumulation into larger slag inclusions. Predictions of final dispersoid size and slag formation using this dissolution zone model match the present experimental data and explain process-structure linkages speculated in the open literature.

36 MATERIALS SCIENCE↗

Space Communications Emulation Facility

Establishing space communication between ground facilities and other satellites is a painstaking task that requires many precise calculations dealing with relay time, atmospheric conditions, and satellite positions, to name a few. The Space Communications Emulation Facility (SCEF) team here at NASA is developing a facility that will approximately emulate the conditions in space that impact space communication. The emulation facility is comprised of a 32 node distributed cluster of computers; each node representing a satellite or ground station. The objective of the satellites is to observe the topography of the Earth (water, vegetation, land, and ice) and relay this information back to the ground stations. Software originally designed by the University of Kansas, labeled the Emulation Manager, controls the interaction of the satellites and ground stations, as well as handling the recording of data. The Emulation Manager is installed on a Linux Operating System, employing both Java and C++ programming codes. The emulation scenarios are written in extensible Markup Language, XML. XML documents are designed to store, carry, and exchange data. With XML documents data can be exchanged between incompatible systems, which makes it ideal for this project because Linux, MAC and Windows Operating Systems are all used. Unfortunately, XML documents cannot display data like HTML documents. Therefore, the SCEF team uses XML Schema Definition (XSD) or just schema to describe the structure of an XML document. Schemas are very important because they have the capability to validate the correctness of data, define restrictions on data, define data formats, and convert data between different data types, among other things. At this time, in order for the Emulation Manager to open and run an XML emulation scenario file, the user must first establish a link between the schema file and the directory under which the XML scenario files are saved. This procedure takes place on the command line on the Linux Operating System. Once this link has been established the Emulation manager validates all the XML files in that directory against the schema file, before the actual scenario is run. Using some very sophisticated commercial software called the Satellite Tool Kit (STK) installed on the Linux box, the Emulation Manager is able to display the data and graphics generated by the execution of a XML emulation scenario file. The Emulation Manager software is written in JAVA programming code. Since the SCEF project is in the developmental stage, the source code for this type of software is being modified to better fit the requirements of the SCEF project. Some parameters for the emulation are hard coded, set at fixed values. Members of the SCEF team are altering the code to allow the user to choose the values of these hard coded parameters by inserting a toolbar onto the preexisting GUI.

Hill, Chante A.↗

Platform for Postprocessing Waveform-Based NDE

Taking advantage of the similarities that exist among all waveform-based non-destructive evaluation (NDE) methods, a common software platform has been developed containing multiple- signal and image-processing techniques for waveforms and images. The NASA NDE Signal and Image Processing software has been developed using the latest versions of LabVIEW, and its associated Advanced Signal Processing and Vision Toolkits. The software is useable on a PC with Windows XP and Windows Vista. The software has been designed with a commercial grade interface in which two main windows, Waveform Window and Image Window, are displayed if the user chooses a waveform file to display. Within these two main windows, most actions are chosen through logically conceived run-time menus. The Waveform Window has plots for both the raw time-domain waves and their frequency- domain transformations (fast Fourier transform and power spectral density). The Image Window shows the C-scan image formed from information of the time-domain waveform (such as peak amplitude) or its frequency-domain transformation at each scan location. The user also has the ability to open an image, or series of images, or a simple set of X-Y paired data set in text format. Each of the Waveform and Image Windows contains menus from which to perform many user actions. An option exists to use raw waves obtained directly from scan, or waves after deconvolution if system wave response is provided. Two types of deconvolution, time-based subtraction or inverse-filter, can be performed to arrive at a deconvolved wave set. Additionally, the menu on the Waveform Window allows preprocessing of waveforms prior to image formation, scaling and display of waveforms, formation of different types of images (including non-standard types such as velocity), gating of portions of waves prior to image formation, and several other miscellaneous and specialized operations. The menu available on the Image Window allows many further image processing and analysis operations, some of which are found in commercially-available image-processing software programs (such as Adobe Photoshop), and some that are not (removing outliers, Bscan information, region-of-interest analysis, line profiles, and precision feature measurements).

Roth, Don↗

Flux Cancelation: The Key to Solar Eruptions

Solar coronal jets are magnetically channeled eruptions that occur in all types of solar environments (e.g. active regions, quiet-Sun regions and coronal holes). Recent studies show that coronal jets are driven by the eruption of small-scare filaments (minifilaments). Once the eruption is underway magnetic reconnection evidently makes the jet spire and the bright emission in the jet base. However, the triggering mechanism of these eruptions and the formation mechanism of the pre-jet minifilaments are still open questions. In this talk, mainly using SDOAIA and SDOHIM data, first I will address the question: what triggers the jet-driving minifilament eruptions in different solar environments (coronal holes, quiet regions, active regions)? Then I will talk about the magnetic field evolution that produces the pre-jet minifilaments. By examining pre-jet evolutionary changes in line-of-sight HMI magnetograms while examining concurrent EUV images of coronal and transition-region emission, we find clear evidence that flux cancelation is the main process that builds pre-jet minifilaments, and is also the main process that triggers the eruptions. I will also present results from our ongoing work indicating that jet-driving minifilament eruptions are analogous to larger-scare filament eruptions that make flares and CMEs. We find that persistent flux cancellation at the neutral line of large-scale filaments often triggers their eruptions. From our observations we infer that flux cancelation is the fundamental process from the buildup and triggering of solar eruptions of all sizes.

Panesar, Navdeep K.↗

Flux Cancelation: The Key to Solar Eruptions

Solar coronal jets are magnetically channeled eruptions that occur in all types of solar environments (e.g. active regions, quiet-Sun regions and coronal holes). Recent studies show that coronal jets are driven by the eruption of small-scale filaments (minifilaments). Once the eruption is underway magnetic reconnection evidently makes the jet spire and the bright emission in the jet base. However, the triggering mechanism of these eruptions and the formation mechanism of the pre-jet minifilaments are still open questions. In this talk, mainly using SDO/AIA and SDO/HMI data, first I will address the question: what triggers the jet-driving minifilament eruptions in different solar environments (coronal holes, quiet regions, active regions)? Then I will talk about the magnetic field evolution that produces the pre-jet minifilaments. By examining pre-jet evolutionary changes in line-of-sight HMI magnetograms while examining concurrent EUV images of coronal and transition-region emission, we find clear evidence that flux cancellation is the main process that builds pre-jet minifilaments, and is also the main process that triggers the eruptions. I will also present results from our ongoing work indicating that jet-driving minifilament eruptions are analogous to larger-scale filament eruptions that make flares and CMEs. We find that persistent flux cancellation at the neutral line of large-scale filaments often triggers their eruptions. From our observations we infer that flux cancellation is the fundamental process for the buildup and triggering of solar eruptions of all sizes.

Panesar, Navdeep K.↗

Flux Cancelation: The Key to Solar Eruptions

Solar coronal jets are magnetically channeled eruptions that occur in all types of solar environments (e.g. active regions, quiet-Sun regions and coronal holes). Recent studies show that coronal jets are driven by the eruption of small-scare filaments (minifilaments). Once the eruption is underway magnetic reconnection evidently makes the jet spire and the bright emission in the jet base. However, the triggering mechanism of these eruptions and the formation mechanism of the pre-jet minifilaments are still open questions. In this talk, mainly using SDOAIA and SDOHIM data, first I will address the question: what triggers the jet-driving minifilament eruptions in different solar environments (coronal holes, quiet regions, active regions)? Then I will talk about the magnetic field evolution that produces the pre-jet minifilaments. By examining pre-jet evolutionary changes in line-of-sight HMI magnetograms while examining concurrent EUV images of coronal and transition-region emission, we find clear evidence that flux cancelation is the main process that builds pre-jet minifilaments, and is also the main process that triggers the eruptions. I will also present results from our ongoing work indicating that jet-driving minifilament eruptions are analogous to larger-scare filament eruptions that make flares and CMEs. We find that persistent flux cancellation at the neutral line of large-scale filaments often triggers their eruptions. From our observations we infer that flux cancelation is the fundamental process from the buildup and triggering of solar eruptions of all sizes.

Panesar, Navdeep K.↗

High-Resolution Gridded Level 3 Aerosol Optical Depth Data from MODIS

The state-of-art satellite observations of atmospheric aerosols over the last two decades from NASA's MODIS instruments have been extensively utilized in climate change and air quality research and applications. The operational algorithms now produce level 2 aerosol data at varying spatial resolutions (1, 3, and 10 km) and level 3 data at 1 degree. The local and global applications have been benefited from the coarse resolution gridded data sets (i.e., level 3, 1 degree), as it is easier to use since data volume is low and, several online and offline tools are readily available to access and analyze the data with minimal computing resources. At the same time, researchers who require data at much finer spatial scales have to go through a challenging process of obtaining, processing, and analyzing larger volumes of data sets that require high-end computing resources and coding skills. Therefore, we have created a high spatial resolution (HRG, 0.1x0.1 degree) daily and monthly aerosol optical depth (AOD) product by combining two MODIS operational algorithms, namely Deep Blue (DB) and Dark Target (DT). The new HRG AODs meets the accuracy requirements of level 2 AOD data and provide either the same or more spatial coverage on daily and monthly scales. The data sets are provided in daily and monthly files through open Ftp server with python scripts to read and map the data. The reduced data volume with an easy to use format and tools to access the data will encourage more users to utilize the data for research and applications.

aerosol↗

An open-access database and analysis tool for perovskite solar cells based on the FAIR data principles

Large datasets are now ubiquitous as technology enables higher-throughput experiments, but rarely can a research field truly benefit from the research data generated due to inconsistent formatting, undocumented storage or improper dissemination. Here we extract all the meaningful device data from peer-reviewed papers on metal-halide perovskite solar cells published so far and make them available in a database. We collect data from over 42,400 photovoltaic devices with up to 100 parameters per device. We then develop open-source and accessible procedures to analyse the data, providing examples of insights that can be gleaned from the analysis of a large dataset. The database, graphics and analysis tools are made available to the community and will continue to evolve as an open-source initiative. This approach of extensively capturing the progress of an entire field, including sorting, interactive exploration and graphical representation of the data, will be applicable to many fields in materials science, engineering and biosciences.

14 SOLAR ENERGY↗

The kinematics of dense clusters of galaxies. 3: Comparison with cosmological models

We compare the combined distribution of 31 group and 25 cluster velocity dispersions with the ensemble of 32 models for the formation and evolution of large-scale structure examined by Weinberg & Cole (1992). The models include Gaussian and non-Gaussian initial fluctuations, different power law spectra (n = -1, n = 0, n = -2, 'pancake'), flat (Omega = 1) and open (Omega = 0.2) cosmologies, and unbiased (b(sub 8) = 1) and biased (b(sub 8) = 2) galaxy formation. The set of initial conditions we test, although limited, samples enough parameter space to indicate which general classes of models are consistent with the data. The two Gaussian, n = -1 models which best approximate the standard and open Cold Dark Matter (CDM) models do not match the observed distribution of velocity dispersions; models with b(sub 8) = 2 and Omega = 1 ('standard') or b(sub 8) = 1 and Omega = 0.2 ('open') predict too large a ratio of low to high velocity dispersion systems. A 'COBE-normalized' CDM model with b(sub 8) = 1 and Omega = 1 produces clusters with velocity dispersions higher than those measured. All three models overestimate the total abundance of systems.

Zabludoff, Ann I.↗

A Processing and Analytics System for Microscopy Data Workflows: The Pycroscopy Ecosystem of Packages

Major advancements in fields as diverse as biology and quantum computing have relied on a multitude of microscopy techniques. Despite the considerable proliferation of these instruments, significant bottlenecks remain in terms of processing, analysis, storage, and retrieval of the acquired datasets. Aside from lack of file standards, individual domain-specific analysis packages are often disjoint from the underlying datasets, and thus keeping track of analysis and processing steps remains tedious for the end-user, hampering reproducibility. Here, in this study, the pycroscopy ecosystem of packages is introduced, an open-source python-based ecosystem underpinned by a common data model. The data model, termed the N-dimensional spectral imaging data format, is realized in pycroscopy's sidpy package. This package is built on top of dask arrays, thus leveraging dask array attributes, but expanding them to accelerate microscopy relevant analysis and visualization. Several examples of the use of the pycroscopy ecosystem to create workflows for data ingestion and analysis of scanning transmission electron microscopy (STEM) and scanning probe microscopy data are shown. Adoption of such standardized routines will be critical to usher in the next generation of autonomous instruments where processing, computation, and meta-data storage will be critical to overall experimental operations.

97 MATHEMATICS AND COMPUTING↗

Surface Complexation/Ion Exchange Hybrid Model for Radionuclide Sorption to Clay Minerals (M4SF-23LL010301062)

This progress report (Level 4 Milestone Number M4SF-23LL010301062) summarizes research conducted at Lawrence Livermore National Laboratory (LLNL) within the Argillite International Collaborations Activity Number SF-23LL01030106. The activity is focused on our long-term commitment to engaging our partners in international nuclear waste repository research. The focus of this milestone is the establishment of international collaborations for surface complexation modeling and the associated impacts of unlocking larger, community-based datasets. More specifically, we are developing a database framework for Spent Fuel and Waste and Science Technology (SFWST) that is aligned with the Helmholtz Zentrum Dresden Rossendorf (HZDR) sorption database development group in support of the database needs of the SFWST program. In our FY22 effort, we described a detailed analysis of U(VI) sorption to quartz through both traditional surface complexation modeling and through a hybrid ML framework. In FY23, effort was placed on publication of these results and expansion of the LLNL surface complexation and ion exchange database (L-SCIE) in order to assess mineral-based radionuclide retardation under a wider variety of geochemical conditions (e.g., ionic strength, varying electrolyte compositions). Efforts were initiated to expand L-SCIE to include radionuclide surface complexation and ion exchange to clays that are relevant to subsurface geochemical processes occurring at nuclear waste repositories. In particular, a large source of sorption data for clays resides at the Paul Scherrer Institute (PSI) (work primarily by Bradbury and Baeyens) and we initiated discussions on how to retrieve those data and apply FAIR principles to those datasets. In addition to L-SCIE development, two hybrid models that incorporate AI/ML were investigated and compared to discern the most promising approaches for accurate and precise estimations of radionuclide retardation. Key considerations for future model development include (1) the ability to reduce computational burden on determining retardation coefficients for PA and (2) the ability to quantify and predict radionuclide-mineral partitioning at a more efficient, rapid pace due to automated workflows. Upon the careful consideration of the most effective modeling approaches, we are identifying ways to implement these approaches into PA. Ultimately, the data science-based workflows will provide a major incentive for other institutions to adopt a FAIR-formatted, interoperable database. LLNL will play a key role in disseminating sorption data and acting as good data stewards by updating the database in a consistent format and assessing the quality of the newly assimilated data in an organized fashion. To this end, all data and workflows are open access and made available on the LLNL Seaborg research website (https://seaborg.llnl.gov/resources/geochemical-databases-modeling-codes).

38 RADIATION CHEMISTRY, RADIOCHEMISTRY, AND NUCLEA↗

Gamma rays and the case for baryon symmetric big-bang cosmology

The baryon symmetric big-bang cosmologies offer an explanation of the present photon-baryon ratio in the universe, the best present explanation of the diffuse gamma-ray background spectrum in the 1 to 200 MeV range, and a mechanism for galaxy formation. In the context of an open universe model, the value of omega which best fits the present gamma-ray data is omega equals approx. 0.1 which does not conflict with upper limits on Comptonization distortion of the 3K background radiation. In regard to He production, evidence is discussed that nucleosynthesis of He may have taken place after the galaxies were formed.

Stecker, F. W.↗

Crystal structures reveal catalytic and regulatory mechanisms of the dual-specificity ubiquitin/FAT10 E1 enzyme Uba6

The E1 enzyme Uba6 initiates signal transduction by activating ubiquitin and the ubiquitin-like protein FAT10 in a two-step process involving sequential catalysis of adenylation and thioester bond formation. To gain mechanistic insights into these processes, we determined the crystal structure of a human Uba6/ubiquitin complex. Two distinct architectures of the complex are observed: one in which Uba6 adopts an open conformation with the active site configured for catalysis of adenylation, and a second drastically different closed conformation in which the adenylation active site is disassembled and reconfigured for catalysis of thioester bond formation. Surprisingly, an inositol hexakisphosphate (InsP6) molecule binds to a previously unidentified allosteric site on Uba6. Our structural, biochemical, and biophysical data indicate that InsP6 allosterically inhibits Uba6 activity by altering interconversion of the open and closed conformations of Uba6 while also enhancing its stability. In addition to revealing the molecular mechanisms of catalysis by Uba6 and allosteric regulation of its activities, our structures provide a framework for developing Uba6-specific inhibitors and raise the possibility of allosteric regulation of other E1s by naturally occurring cellular metabolites.

59 BASIC BIOLOGICAL SCIENCES↗

COMPASS-FME Synoptic Sites Level 2 Sensor Data v2-1

This is the version 2-1 Level 2 (L2) data release for COMPASS-FME environmental sensors located at our synoptic field sites. COMPASS-FME is studying sites in two distinct regions, the Chesapeake Bay and the Western Lake Erie Basin. We established the network at seven "synoptic" (observational) sites along the Chesapeake Bay and Lake Erie coastlines, collectively generating over three million observations per month, to track and comprehend environmental changes where land and water intersect. Additionally, the two regions provide an interesting contrast of saltwater and freshwater coasts that allow us to differentiate the impacts of inundation and coastal water chemistries in two nationally important coastal systems. Level 2 (L2) data consist of sensor observations from the COMPASS-FME synoptic sites, TEMPEST, and DELUGE. Compared to the L1 data, these are more consistent (always 15-minute timestamps for the entire year); better QA/QC’d (out of bounds, out of service, and extreme outlier values are removed); and more complete, with a gap-filled time series available alongside the main observations, and additional derived (calculated) variables. L2 data are intended to be rapidly and easily usable in analyses and simulations. However, algorithmic outlier identification always carries the risk of removing valid data, and Level 1 data may be more suitable for analyses that focus on variability or extreme events. This dataset includes: - An overall dataset README file that describes the current version, gives citation and contact information, etc. - Site- and year-specific folders, each holding variable-specific Parquet (a high performance, space efficient format; see https://parquet.apache.org) data files for each site and plot in that year. - Metadata files within each site-year folder provide full information on data units, expected ranges, contact information, detailed flood times, as well as a general description of the site. - Environmental sensor types that appear in the data files include weather (ClimaVUE50, CS, RM Young, and LI instruments in the graphs below); soil conditions (TEROS12); soil redox state (Redox); groundwater variables (AquaTROLL200 and AquaTROLL600); open water sondes (Exo); tree sap velocity (Sapflow); and system voltage and state (Datalogger). Data are reported every 15 minutes. Data files are in Apache Parquet, a high performance, space efficient format for tabular data. These files can be read using R's `arrow` package (https://arrow.apache.org/docs/r/), with similar tools available in other languages. Please see v2-1 L2 Sensor Package QStart.pdf for detailed information on data package structure, temporal coverage, and versioning.

EARTH SCIENCE > ATMOSPHERE > ATMOSPHERIC TEMPERATU↗

HAPI: An API Standard for Accessing Heliophysics Time Series Data

Heliophysics data analysis often involves combining diverse science measurements, many of them captured as time series. Although there are now only a few commonly used data file formats, the diversity in mechanisms for automated access to and aggregation of such data holdings can make analysis that requires intercomparison of data from multiple data providers difficult. The Heliophysics Application Programmer's Interface (HAPI) is a recently developed standard for accessing distributed time series data to increase interoperability. The HAPI specification is based on the common elements of existing data services, and it standardizes the two main parts of a data service: the request interface and the response data structures. The interface is based on the REpresentational State Transfer (REST) or RESTful architecture style, and the HAPI specification defines five required REST endpoints. Data are returned via a streaming format that hides file boundaries; the metadata is detailed enough for the content to be scientifically useful, e.g., plotted with appropriate axes layout, units, and labels. Multiple mature HAPI-related open-source projects offer server-side implementation tools and client-side libraries for reading HAPI data in multiple languages (IDL, Java, MATLAB, and Python). Multiple data providers in the US and Europe have added HAPI access alongside their existing interfaces. Based on this experience, data can be served via HAPI with little or no information loss compared to similar existing web interfaces. Finally, HAPI has been recommended as a COSPAR standard for time series data delivery.

Robert S. Weigel↗

Generalizable Web User Interface for Scalable and Streamlined Deployment of Building Energy Management Systems in Small and Medium-Sized Commercial Buildings

Small and medium-sized commercial buildings (SMCBs) comprise 94% of US commercial buildings yet face significant barriers to implementing building energy management systems despite advances in smart device technology. Existing solutions present critical limitations: cloud-based API solutions simplify deployment but create vendor lock-in constraints; commercial integrated software solutions ensure compatibility via standardized protocols but require substantial cost and technical expertise; open-source IoT platforms offer cost-effective vendor independence but provide insufficient standardized protocol support for commercial building automation. This research presents a generalizable web user interface framework that bridges the gap between evolving smart device capabilities and lagging software infrastructure for SMCBs. The proposed system integrates VOLTTRON open-source middleware with an automated configuration converter that transforms unified specifications written in YAML, a human-readable data-serialization format, into system-specific files, streamlining manual setup processes. The vendor-agnostic architecture supports industry-standard protocols (BACnet and Modbus) and semantic building models while providing adaptive web interfaces that dynamically adjust to various building configurations. Demonstrations through simulation-based testing and a field deployment show automatic interface adaptation across heterogeneous HVAC systems and multizone monitoring. The automated configuration converter also substantially reduces labor-intensive setup.

Chung, Jihoon [ORNL] (ORCID:0000000184880815)↗

Symphony: Cosmological Zoom-in Simulation Suites over Four Decades of Host Halo Mass

Abstract We present Symphony, a compilation of 262 cosmological, cold-dark-matter-only zoom-in simulations spanning four decades of host halo mass, from 10 11 –10 15 M ⊙ . This compilation includes three existing simulation suites at the cluster and Milky Way–mass scales, and two new suites: 39 Large Magellanic Cloud-mass (10 11 M ⊙ ) and 49 strong-lens-analog (10 13 M ⊙ ) group-mass hosts. Across the entire host halo mass range, the highest-resolution regions in these simulations are resolved with a dark matter particle mass of ≈3 × 10 −7 times the host virial mass and a Plummer-equivalent gravitational softening length of ≈9 × 10 −4 times the host virial radius, on average. We measure correlations between subhalo abundance and host concentration, formation time, and maximum subhalo mass, all of which peak at the Milky Way host halo mass scale. Subhalo abundances are ≈50% higher in clusters than in lower-mass hosts at fixed sub-to-host halo mass ratios. Subhalo radial distributions are approximately self-similar as a function of host mass and are less concentrated than hosts’ underlying dark matter distributions. We compare our results to the semianalytic model Galacticus , which predicts subhalo mass functions with a higher normalization at the low-mass end and radial distributions that are slightly more concentrated than Symphony. We use UniverseMachine to model halo and subhalo star formation histories in Symphony, and we demonstrate that these predictions resolve the formation histories of the halos that host nearly all currently observable satellite galaxies in the universe. To promote open use of Symphony, data products are publicly available at http://web.stanford.edu/group/gfc/symphony .

79 ASTRONOMY AND ASTROPHYSICS↗

Closing the Gap between FAIR Data Repositories and Hierarchical Data Formats

Many in the scientific community, particularly in publicly funded research, are pushing to adhere to more accessible data standards to maximize the findability, accessibility, interoperability, and reusability (FAIR) of scientific data, especially with the growing prevalence of machine learning augmented research. Online FAIR data repositories, such as the Open Science Framework (OSF), help facilitate the adoption of these standards by providing frameworks for storage, access, search, APIs, and other features that create organized hubs of scientific data. However, the wider acceptance of such repositories is hindered by the lack of support of hierarchical data formats, such as Technical Data Management Streaming (TDMS) and Hierarchical Data Format 5 (HDF5), that many researchers rely on to organize their datasets. Various tools and strategies should be used to allow hierarchical data formats, FAIR data repositories, and scientific organizations to work more seamlessly together. A pilot project at Los Alamos National Laboratory (LANL) addresses the disconnect between them by integrating the OSF FAIR data repository with hierarchical data renderers, extending support for additional file types in their framework. The multifaceted interactive renderer displays a tree of metadata alongside a table and plot of the data channels in the file. This allows users to quickly and efficiently load large and complex data files directly in the OSF webapp. Users who are browsing files can quickly and intuitively see the files in the way they or their colleagues structured the hierarchical form and immediately grasp their contents. This solution helps bridge the gap between hierarchical data storage techniques and FAIR data repositories, making both of them more viable options for scientific institutions like LANL which have been put off by the lack of integration between them.

97 MATHEMATICS AND COMPUTING↗