FlowBench Raw Data Archive
The repo that provides data archive for DOE PoSeiDon project. It also contains scripts and instructions to parse the data.
SEARCH · Engineering Papers
Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.
Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.
The repo that provides data archive for DOE PoSeiDon project. It also contains scripts and instructions to parse the data.
The adequate simulation of internal climate variability is key for our understanding of climate as it underpins efforts to attribute historical events, predict on seasonal and decadal time scales, and isolate the effects of climate change. Here the skill of models in reproducing observed modes of climate variability is assessed, both across and within the CMIP3, CMIP5, and CMIP6 archives, in order to document model capabilities, progress across ensembles, and persisting biases. A focus is given to the well-observed tropical and extratropical modes that exhibit small intrinsic variability relative to model structural uncertainty. These include El Niño–Southern Oscillation (ENSO), the Pacific decadal oscillation (PDO), the North Atlantic Oscillation (NAO), and the northern and southern annular modes (NAM and SAM). Significant improvements are identified in models’ representation of many modes. Canonical biases, which involve both amplitudes and patterns, are generally reduced across model generations. For example, biases in ENSO-related equatorial Pacific sea surface temperature, which extend too far westward, and associated atmospheric teleconnections, which are too weak, are reduced. Stronger tropical expression of the PDO in successive CMIP generations has characterized their improvement, with some CMIP6 models generating patterns that lie within the range of observed estimates. For the NAO, NAM, and SAM, pattern correlations with observations are generally higher than for other modes and slight improvements are identified across successive model generations. Finally, for ENSO and PDO spectra and extratropical modes, changes are small compared to internal variability, precluding definitive statements regarding improvement.
This Modeling Archive supports the ORNL-SFA and IDEAS-Waterhsed publication submitted to Environmental Modeling and Software. A recently introduced multiscale model for representing the combined effect of hyporheic exchange flows and small-scale hyporheic-zone biogeochemical processes is extended from the reach scale to river network scales. The ADELS (Advection Dispersion Equation with Lagrangian Subgrid) model uses a one-dimensional advection-dispersion-reaction equation for the channel network and couples that equation at each channel location to a one-dimensional advection-reaction subgrid model representing an ensemble of streamlines that are diverted into the hyporheic zone before returning to the channel. The subgrid model is written in a computationally advantageous Lagrangian form with hyporheic age replacing the hyporheic travel distance. In the paper, we summarized implementation in the integrated surface/subsurface hydrology modeling system Advanced Terrestrial Simulator (ATS).
Summit Darshan Archival Dataset contains 2021 Summit Darshan log data for 25 applications and is grouped into science domains. The dataset is processed, and all the propriety fields are anonymized. The resultant data is converted into a tabular structure and saved in parquet file format. In this notebook, we demonstrate how to access the data. Data Organization: The data is organized into two directories: Darshan total (`darshan_total`): List all the high levels generated by the `darshan-parser --total` command on `.darshan` files. There is one parquet file for each application. Note: `uid` and `exe` field are masked Darshan detail (`darshan_detail`): This data contains detailed job level log information extracted by command `darshan-parser` on the raw `.darshan` files. The data is sorted by directory hierarchy in the order of `year/month/day (2021/12/07)`. For instance, to get the data for a `job_id` 3819766 of application `App11`, which was executed on `2021-12-07`can be accessed as follows. Note:`uid` and `filename` fields are masked
This is the data archive for manuscript PPCF-104238 at Plasma Physics and Controlled Fusion. Authors: Ligia D. Amorim, Carlo Benedetti, Stepan S. Bulanov, Davide Terzani, Axel Huebl, Carl B. Schroeder, Jean-Luc Vay, Eric Esarey
This dataset contains input, output, and sensitivity data files for computational simulations with the SCALE code system as part of the Verified, Archived Library of Inputs and Data (VALID). The simulations cover critical benchmark experiments from the International Criticality Safety Benchmark Evaluation Project. The files are to be housed in a public directory for distribution. The information contained in the files have been approved for release by the Organisation for Economic Co-operation and Development Nuclear Energy Agency (NEA). Users wanting to reproduce results from this dataset are required to obtain a license to the SCALE code system for which details on the distribution can be found here: https://www.ornl.gov/scale/releases.
This archive contains the data and Python scripts required to reproduce the analyses and figures in the study: Gomez-Velez, J. D., Rathore, S. S., Cohen, M. J., & Painter, S. L. (2025). Hyporheic-zone Processes and Stream Oxygen Dynamics: Insights from a Multiscale Reactive Transport Model. Submitted to Water Resources Research. The analysis utilizes the subgrid model Advection Dispersion Equation with Lagrangian Subgrids (ADELS) implemented in the Advanced Terrestrial Simulator (ATS; https://amanzi.github.io/ats/stable/). In this case, the ATS and Amanzi versions are (1) ATS version 1.5.1_f5ba18f8 and (2) Amanzi version 1.6-dev_53444cca4. The repository includes a Jupyter Notebook and the necessary data (Pandas DataFrames stored as pickle files) to generate the figures for the manuscript. Additionally, it contains Python scripts to create ATS input files, run the ATS simulations, and post-process the results. Finally, it provides routines for parameter estimation using the Single-Station Metabolism (SSM) model with the Differential Evolution Adaptive Metropolis (DREAM) Markov Chain Monte Carlo (MCMC) algorithm with ZS enhancements (DREAM-ZS).
This dataset is a model archive of the paper Permafrost Thaw, Uneven Subsidence and Projected Drying of Ice-wedge Polygon Tundra (in prep) to support a modeling study investigating how projected increases in Arctic temperature and precipitation will jointly influence hydrologic conditions in ice-rich tundra landscapes. With this dataset, this study is to address the research question: Will Arctic tundra landscapes become wetter or drier with increasing precipitation and temperature in the future when thaw-induced ground subsidence and associated microtopographic evolution are represented? The simulations focus on ice-wedge polygon tundra, a widespread form of ice-rich permafrost terrain that is highly sensitive to thaw-driven landscape change. This dataset contains model input and output data for four study watersheds in Alaska: Anaktuvuk, Utqiagvik (formerly Barrow), Brooks Foothills, and Prudhoe Bay. Simulations were performed using the Advanced Terrestrial Simulator (ATS, v1.5), a physics-rich integrated surface–subsurface hydrologic model. For each watershed, ten modeling cases were performed representing two landscape evolution conditions (with subsidence and without subsidence) combined with five climate forcing scenarios derived from Shared Socioeconomic Pathways (SSP5, SSP5 with precipitation trend, SSP2, SSP2 with precipitation trend, and SSP2 with double precipitation trend). Particularly, for each watershed under the forcing SSP2 with precipitation trend, there are two additional simulations considering spatially heterogeneous subsidence distributions: one assumes randomly distributed scaling and the other includes elevation dependent distribution scaling. These simulations span 1980–2099 and include spin-up runs (1980–2009) followed by transient projections (2010–2099). To facilitate reproducibility of simulations, all datasets are organized by watershed. For each study watershed, the dataset contains: (1) Pre-partitioned mesh files for 32-core modeling (.par.32.XX), located in EACH_WATERSHED/mesh/basin; and also a non-partitioned mesh file (.exo) located in EACH_WATERSHED/mesh; (2) Climate forcings corresponding to the five SSP scenarios (.h5), located in EACH_WATERSHED/data; (3) Final states (.h5) from column spin-up modeling used to initialize historical watershed-scale spin-up runs from 1980 to 2009, located in EACH_WATERSHED/PreSpinupHistorical; (4) Final states (.h5) of historical watershed-scale spin-up runs from 1980 to 2009 used to initialize projection runs, located in EACH_WATERSHED/Spinup_daymetERA5; (5) ATS modeling input files (.xml), located in EACH_WATERSHED/EACH_SIMULATION_SCENARIO/inputfiles; (6) ATS modeling output files (.dat), located in in EACH_WATERSHED/EACH_SIMULATION_SCENARIO/combined_obs; (7) For the Brooks Foothills watershed, additional spatial model outputs are provided (.h5) for selected years (2033 and 2093) used to generate spatial figures in this study, located in Brooksfoothills/EACH_SIMULATION_SCENARIO/results-WITH/WITHOUT_SUBSIDENCE-year2033/2093. All data files with suffix .h5 can be accessible through Python h5py, and all data files with suffix of .dat can be imported by Python pandas. Mesh file with .exo can be visualized through Paraview or read by Python netCDF. The Next-Generation Ecosystem Experiments in the Arctic (NGEE Arctic) project is a research effort to reduce uncertainty in the Department of Energy’s Energy Exascale Earth System Model (E3SM) by developing a predictive understanding of Arctic tundra ecosystems underlain by permafrost and to quantify feedbacks from the Arctic tundra to the Earth system. NGEE Arctic is supported by the Department of Energy's Office of Biological and Environmental Research. Over Phases 1–3, observations made by the NGEE Arctic team across a gradient of permafrost landscapes in Arctic Alaska improved the representation of tundra processes in the land surface component of E3SM (the E3SM Land Model, ELM). Model improvements emphasized unique aspects of permafrost environments and explored reductions in model complexity while retaining predictive power. The Arctic-informed ELM developed by NGEE Arctic has been used to make novel predictions on processes ranging from permafrost thaw to soil biogeochemical cycling to Earth system feedbacks associated with the unique characteristics of tundra plants. In Phase 4, the NGEE Arctic team is evaluating our new predictive understanding under novel conditions across the Arctic domain. In collaboration with partners at long-term pan-Arctic research sites we are examining whether an Arctic-informed ELM can faithfully simulate interactions among surface and subsurface processes at site, regional, and pan-Arctic scales. In turn, we are using variety of tools to dynamically extend and evaluate ELM inference, with an emphasis on data synthesis and pan-Arctic model evaluation, reintegration of code with an evolving E3SM, scaling across heterogeneous Arctic landscapes, and the appropriate representation of the impacts of increasingly frequent Arctic disturbances.
This dataset is a model archive of the paper A bespoke model of Arctic river basins based on hillslope delineation (in prep), which introduces a watershed decomposition and parameterization method for large scale permafrost hydrology simulation. With this dataset, this study aims to address the research question: whether a computationally efficient hillslope-based modeling framework can reliably simulate discharge at Arctic river-basin scales. This dataset contains model input and output data for five modeling scenarios at a study site located in the Sagavanirktok River basin. The five modeling scenarios include three modeling cases under temperate conditions using full 3D, decomposed 3D, and decomposed 2D modeling strategies; and two modeling cases under actual Arctic conditions with permafrost using full 3D and decomposed 2D modeling strategies. Simulations were performed using the Advanced Terrestrial Simulator (ATS, v1.6 for three temperate scenarios and v1.5 for two Arctic scenarios), a physics-rich integrated surface–subsurface hydrologic model with cryo-hydrology features. For the three temperate models, simulations were conducted for the period of 10/01/1993 - 09/30/2002; and for the two Arctic models, simulations were conducted for the period of 01/01/1994 - 12/31/2002. To facilitate reproducibility of simulations, all datasets are organized hierarchically. The dataset contains: (1) Mesh files (.exo) for full 3D model, decomposed 3D models, and decomposed 2D models, located in huc/190604020802_gauge15906000/mesh/. Mesh files can be visualized through Paraview or read by Python. (2) Climate forcings (.h5) for full 3D model and decomposed 3D/2D models are located in huc/190604020802_gauge15906000/daymet_onePiece/, and huc/190604020802_gauge15906000/vp_pr_revised_daymet_1980_2006_with_wind/ separately. Accessible by Python. (3) Raw measured gage discharge (.csv) from USGS, located in huc/190604020802_gauge15906000/gaged_basin15906000_discharge_usgs/. Accessible by Python. (4) Delineated subdomain raster (.tif) and shape files (.shp), and the final parameterized results (.npy) for decomposed models, located in huc/190604020802_gauge15906000/data_preprocessed-meshing. Accessible by Python. (5) Temperate models are located in nonpermaf_huc190604020802_gauge15906000/, which includes three cases: decomposed 2D models (inside model_0*-hillslope_*), decomposed 3D models (inside model_1*-subcatchment_*), and full 3D model (inside model_2*-onepiece_*). Two step spin-up results (checkpoint_final.h5) are located in model_*1-*_spinup_steadystate and model_*2-*_spinup_cycle, separately, which are used to initialize real transient models. The input files (.xml) and output results (.dat) of the real transient models are located in model_*3-*_transient/. Especially, for two example hillslope models (ID=-11 and 11), additional h5py files are included in model_03-hillslope_transient/hillslope-11/, model_03-hillslope_transient/hillslope11, model_13-subcatchment_transient/subcatchment-11/, model_13-subcatchment_transient/subcatchment/11, respectively, which are used to plot the saturation figure (Figure 5) in the manuscript. Accessible by Python. (6) Arctic models are located in huc190604020802_gauge15906000/, which includes two cases: decomposed 2D models (inside model_04-hillslope_transient), and full 3D model (inside model_05-onepiece_transient_mannp1_ra). Three step spin-up results (checkpoint_final.h5) are located in model_01-column_freezeup/, model_02-column_spinup/, model_03-hillslope_spinup/, respectively, which are used to initialize real 2D transient hillslope models. The input files (.xml) and output results (.dat) of transient 2D hillslope models are located in model_04-hillslope_transient/. The input files (.xml) and output results (.dat) of the full 3D transient model is located in model_05-onepiece_transient_mannp1_ra/. The full 3D transient model is initialized by model_02-column_spinup/. Accessible by Python. (7) The MOSART routed discharge results (.csv) under Arctic conditions is located in huc190604020802_gauge15906000/MOSART/. Accessible by Python. (8) All Python codes (.py) used to parameterize full 3D model to decomposed 2D models are located in script/. These codes fit with watershed workflow (a watershed delineation tool) v1.4 under the branch gaob/v1.4 from https://github.com/gaobhub/watershed-workflow.git.
This paper presents an idea to develop a “Rosetta Stone” for unifying observations from various satellite or remote sensors into a common format that would vastly advance our ability to exploit existing datasets for improving predictability within Earth System Models (ESMs). While the applications of such a unified archive are broad, we believe it will be a critical step toward ushering in a new generation of ESMs that are richly informed, guided by, and validated by extensive observational data. With the vast quantity of both remotely-sensed and in-situ data streams available and coming online, new approaches are needed that can harmonize and thus fully exploit these expensive datasets. While we present the broader idea, we point to examples of applications that impact the water cycle and its representation in ESMs.
BEE will have the ability to archive, clone, and re-run workflows that have been previously executed. This will enable reproducibility and portability of scientific workflows both within and between DOE facilities.
EPOC uses the Deep Dive process to discuss and analyze current and planned science use cases and anticipated data output of a particular use case, site, or project to help inform the strategic planning of a campus or regional networking environment. This includes understanding future needs related to network operations, network capacity upgrades, and other technological service investments. A Deep Dive comprehensively surveys major research stakeholders’ plans and processes in order to investigate data management requirements over the next 5–10 years. Deep Dives help ensure that key stakeholders have a common understanding of the issues and the actions that a campus or regional network may need to undertake to offer solutions. The EPOC team leads the effort and relies on collaboration with the hosting site or network, and other affiliated entities that participate in the process. EPOC organizes, convenes, executes, and shares the outcomes of the review with all stakeholders. Between May 2021 and August 2021, staff members from the Engagement and Performance Operations Center (EPOC) met with researchers and staff from the National Oceanic and Atmospheric Administration (NOAA)'s N-Wave (the Enterprise network that supports the NOAA mission) and National Centers for Environmental Information (NCEI)'s Fisheries Acoustics Archive for the purpose of a recording a Deep Dive into research drivers. The goal of these meetings was to help characterize the requirements for the research use case, and to enable cyberinfrastructure support staff to better understand the needs of the researchers they support.
This report describes the development of a comprehensive catalogue of generic features, events, and processes (FEPs) that are potentially important for the post-closure performance of a repository for high-level radioactive waste (HLW) and spent nuclear fuel (SNF) in salt (halite) host rock. The FEPs and other supporting information have been entered into a “SaltFEP” Database. The generic salt repository FEPs include consideration of relevant FEPs from a number of U.S., Dutch, German, and international FEP lists and should be a suitable starting point for any repository program in salt host rock. The salt FEP catalogue and database employ a FEP classification matrix approach that is based on the concept that a FEP is typically a process or event acting upon or within a feature. The FEP matrix provides a two-dimensional structure consisting of a Features/Components axis that defines the “rows” and a Processes/Events axis that defines the “columns” of the matrix. The design of the FEP classification matrix is consistent with repository performance assessment – the Features/Components axis is organized vertically to generally correspond to the direction of potential radionuclide migration (from the waste to the biosphere) and the Processes/Events axis is designed to represent the common two-way couplings between thermal processes and other processes (such as thermal-mechanical or thermal-hydrologic processes). Related FEPs can be easily identified – related FEPs will typically be grouped in a single matrix cell or aligned along a common row (Feature/Component) or column (Process/Event). The online SaltFEP database can be downloaded from www.saltfep.org. It contains the FEP matrix, the FEPs, and the associated processes for each FEP. It provides a starting point to create and document site-specific individual FEPs. Furthermore, the FEP matrix is connected to the Salt Knowledge Archive (SKA), a database of about 20,000 references and documents representing the historical knowledge on radioactive disposal in salt. This work is the result of an ongoing collaboration between researchers in the U.S., the Netherlands, and Germany, and supports the NEA Salt Club Mandate. It builds upon prior work which is documented.
Y-12 has made significant progress in FY21 towards National Nuclear Materials Archive (NNMA) program objectives related to the identification, nomination, sampling, sub-sampling and shipment of NNMA materials. Y-12 efforts in FY22 and beyond are anticipated to transition away from material identification and nomination to an emphasis on sampling/sub-sampling and long-term storage of NNMA specimens. Long-term storage efforts will also involve more emphasis on insuring that the nuclear forensics value of NNMA items is guaranteed through appropriate preservation processes.
The objective of this project was to create a searchable database of plutonium compatibility studies from the Rocky Flats Archive. The Rocky Flats Plant was a manufacturing complex in Golden, Colorado that produced nuclear weapons. It primarily focused on producing plutonium pits, which were the cores of many nuclear implosion-type weapons. Because plutonium is an extremely reactive metal, avoiding the use of incompatible materials is imperative to avoid damaging the plutonium pit during its production. The Rocky Flats Plant had conducted extensive surveys of materials used in plutonium pit fabrication for their compatibility with plutonium.
This work details the development of a concentrating solar power (CSP) and thermal (CST) library archive. This work included digitization of one-of-a-kind documents that could be degraded or destroyed over time. Sandia National Laboratories (SNL) National Solar Thermal Test Facility (NSTTF) and Sandia's Technical Library departments collaborated to establish and maintain the first and only digital collection in the world of Concentrating Solar Power (CSP) related historical documents. These date back to the CSP program inception here at Sandia in the early 1970's thru to the present.
The scale and speed of data generated by modern scientific experiments have constantly challenged the research community to store, curate, manage and optimally use it to drive scientific discoveries. In this work, we have developed a data archive and portal (DAP) platform including analytics capabilities to collect, curate, and manage data and metadata stream for solid phase processing (SPP) techniques. We successfully hosted around ~347K files of data related to processing parameters, microscopic images, and spectroscopic data related to solid phase processing. The DAP platform for SPP will establish an enduring capability to support machine learning and grow collaboration at the intersection of materials science and data science.
This procedure provides a framework for preparing, reviewing, and storing model inputs and derived data so that individuals with authorized access to the Verified, Archived, Library of Inputs and Data (VALID) repository can use the inputs and data with confidence in their analyses. This procedure uses documented checks and reviews to ensure that the inputs and data were correctly generated using appropriate references. Configuration management is implemented to prevent inadvertent modification of the inputs and data or inclusion of models that have not been reviewed. This procedure also provides guidance to be followed if errors are identified or if input or data revisions are needed.