Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Data models”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

Exploring the Limits of the Data-Model-Theory Synergy: “Hot” MW Transitions for Rovibrational IR Studies

In order to further improve the accuracy of rovibrational IR line lists generated from the “Best Theory +Reliable High-resolution Experiment” (BTRHE) strategy from 0.01-0.05 cm-1, or 300-1500 MHz, to ~10 MHz, we explore the current limits of the Data-Model-Theory synergy by examining the accuracy and consistency of existing data, then propose that “hot” bands in microwave (MW) spectra is the solution we need for future enhancements. The Ames SO2 J=0-20 rovibrational energy levels computed on the semi-empirically refined Ames-2 potential energy surface (PES) are fit to the Effective Hamiltonian (EH) model regularly used in the experimental infrared (IR) analysis for SO2 isotopologues. In the fitted EH(Ames) model, the rotational constants A/B/C and all 5 quartic centrifugal distortion constants display clear, systematic, and consistent patterns along the vibrational state energy or quanta. Such consistent patterns may facilitate the vibrational assignments for MW hot bands and extract more information from high temperature MW spectra. Some EH(Expt) analyses were carried out with the lowest order Coriolis Coupling term, C1. Their constants should not be directly compared with other EH(Expt) and EH(Ames) results. After excluding them, our  = EH(Ames)- EH(Expt) analyses for 5 isotopologues (626, 636, 646, 628 and 828) indicates some loss of accuracy and consistency starting from vibrational states as low as 22 or 1000 cm-1. Some EH parameters, e.g. K, may have relative deviations as large as 50-100% and totally lose any recognizable patterns. This simply means that current EH(Expt) models do not have the system-wide consistency we need to further refine the EH(Ames) and Ames rovibrational IR line lists. A large part of such defects are probably inherited from the limited precision of experimental line positions, i.e. 1E-3 ~ 1E-4 cm-1, or 3-30 MHz. This is confirmed in a series of truncation tests using the Ames data. Although the EH(Ames) consistency may help identify unreliable rovibrational EH(Expt) parameters, and make reliable predictions for minor isotopologues and unobserved vibrational bands, we believe only the highly accurate “hot” MW transitions can provide real enhancements for EH(Expt) accuracy and consistency. “Hot” MW spectra should play a more significant role in the future synergy of Data, Model, and Theory in the field of rovibrational IR studies.

Xinchuan Huang↗

Scalable Volume Visualization for Big Scientific Data Modeled by Functional Approximation

Considering the challenges posed by the space and time complexities in handling extensive scientific volumetric data, various data representations have been developed for the analysis of large-scale scientific data. Multivariate functional approximation (MFA) is an innovative data model designed to tackle substantial challenges in scientific data analysis. It computes values and derivatives with high-order accuracy throughout the spatial domain, mitigating artifacts associated with zero- or first-order interpolation. However, the slow query time through MFA makes it less suitable for interactively visualizing a large MFA model. In this work, we develop the first scalable interactive volume visualization pipeline, MFA-DVV, for the MFA model encoded from large-scale datasets. Our method achieves low input latency through distributed architecture, and its performance can be further enhanced by utilizing a compressed MFA model while still maintaining a high-quality rendering result for scientific datasets. We conduct comprehensive experiments to show that MFA-DVV can decrease the input latency and achieve superior visualization results for big scientific data compared with existing approaches.

big scientific dataset↗

PV Performance Modeling - Data and Resources

The Photovoltaic (PV) Performance Modeling Collaborative (PVPMC) organized a blind PV performance modeling intercomparison to allow PV modelers to blindly test their models and modeling ability against real system data. Measured weather and irradiance data were provided along with detailed descriptions of PV systems from two locations (Albuquerque, New Mexico, USA and Roskilde, Denmark). Participants were asked to simulate the plane-of-array irradiance, module temperature, and DC power output from six systems and submit their results to Sandia for processing. This dataset includes seven MS-Excel sheets with instructions, notes and all necessary data (weather, irradiance, temperature, power) used for the data analysis of the blind modeling comparison. The hourly data represent six different systems from Albuquerque, NM and Roskilde, Denmark over a period of one year. These data are useful for PV performance model validation studies.

14 SOLAR ENERGY↗

Guidelines for Publicly Archiving Terrestrial Model Data to Enhance Usability, Intercomparison, and Synthesis

Scientific communities are increasingly publishing data to evaluate, accredit, and build on published research. However, guidelines for curating data for publication are sparse for model-related research, limiting the usability of archived simulation data. In particular, there are no established guidelines for archiving data related to terrestrial models that simulate land processes and their coupled interactions with climate. Terrestrial modelers have a unique set of challenges when publishing data due to the diversity of scientific domains, research questions, and the types and scales of simulations. Researchers in the U.S. Department of Energy’s (DOE) projects use a variety of multiscale models to advance robust predictions of terrestrial and subsurface ecosystem processes. Here, we synthesize archiving needs for data associated with different DOE models, and provide guidelines for publishing terrestrial model data components following FAIR (Findable, Accessible, Interoperable, Reusable) principles. The guidelines recommend archiving model inputs and testing data used in final simulation runs along with associated codes, workflow scripts, and metadata in public repositories. Researchers should consider archiving model outputs if they are within the storage limits of the repository. We also provide considerations for how to bundle files into different data publications with citable digital object identifiers. Finally, we identify repository features and tools that would enable storage and reuse of model data. Given the diversity of DOE terrestrial models, these guidelines are transferable to other model types and will enable efficient reuse of simulation data for purposes such as model intercomparisons, initialization, benchmarking, synthesis, and comparisons with field observations.

58 GEOSCIENCES↗

An interactive environment for the analysis of large Earth observation and model data sets

We propose to develop an interactive environment for the analysis of large Earth science observation and model data sets. We will use a standard scientific data storage format and a large capacity (greater than 20 GB) optical disk system for data management; develop libraries for coordinate transformation and regridding of data sets; modify the NCSA X Image and X Data Slice software for typical Earth observation data sets by including map transformations and missing data handling; develop analysis tools for common mathematical and statistical operations; integrate the components described above into a system for the analysis and comparison of observations and model results; and distribute software and documentation to the scientific community.

Bowman, Kenneth P.↗

An interactive environment for the analysis of large Earth observation and model data sets

We propose to develop an interactive environment for the analysis of large Earth science observation and model data sets. We will use a standard scientific data storage format and a large capacity (greater than 20 GB) optical disk system for data management; develop libraries for coordinate transformation and regridding of data sets; modify the NCSA X Image and X DataSlice software for typical Earth observation data sets by including map transformations and missing data handling; develop analysis tools for common mathematical and statistical operations; integrate the components described above into a system for the analysis and comparison of observations and model results; and distribute software and documentation to the scientific community.

Bowman, Kenneth P.↗

Computerized design of controllers using data models

The major contributions of the grant effort have been the enhancement of the Compensator Improvement Program (CIP), which resulted in the Ohio University CIP (OUCIP) package, and the development of the Model and Data-Oriented Computer Aided Design System (MADCADS). Incorporation of direct z-domain designs into CIP was tested and determined to be numerically ill-conditioned for the type of lightly damped problems for which the development was intended. Therefore, it was decided to pursue the development of z-plane designs in the w-plane, and to make this conversion transparent to the user. The analytical development needed for this feature, as well as that needed for including compensator damping ratios and DC gain specifications, closed loop stability requirements, and closed loop disturbance rejection specifications into OUCIP are all contained in Section 3. OUCIP was successfully tested with several example systems to verify proper operation of existing and new features. The extension of the CIP philosophy and algorithmic approach to handle modern multivariable controller design criteria was implemented and tested. Several new algorithms for implementing the search approach to modern multivariable control system design were developed and tested. This analytical development, most of which was incorporated into the MADCADS software package, is described in Section 4, which also includes results of the application of MADCADS to the MSFC ACES facility and the Hubble Space Telescope.

Irwin, Dennis↗

Differences Between OCO‐2 and GOME‐2 SIF Products From a Model‐Data Fusion Perspective

Space-borne retrievals of solar-induced chlorophyll fluorescence (SIF) over land surfaces have recently become a resource for studying and quantifying the broad scale dynamics of gross carbon uptake (gross primary productivity—GPP) across ecosystems. To prepare for the assimilation of SIF data in terrestrial biosphere models, we examine how differences between SIF products (due to differences in acquisition characteristics and processing chain) may affect the optimization of model parameters and the resultant GPP estimate. We compare recent daily mean SIF products (one from the Orbiting Carbon Observatory-2 [OCO-2] and two from the Global Ozone Monitoring Experiment–2 [GOME-2], GlobFluo [GF] and NASA-v28 [N28], missions), averaged at 0.5° × 0.5° spatial resolution and 16-day temporal resolution, at the biome level. Phase differences between these products are relatively small. A first-order correction of the difference in spectral sampling between the two instruments shows that OCO-2 and N28 are consistent in terms of magnitude and amplitude, while GF is twice as large as the others. Using a bias-blind toy data assimilation framework, we analyze how biases between SIF products, and between model and products, can be partially alleviated by optimizing the slope and intercept parameters of a linear GPP-SIF operator. As observation biases can transfer to biases in other optimized process-based parameters and to modeled carbon fluxes— thereby resulting in unidentified inaccurate parameter values—we argue that potential SIF biases should be treated cautiously in real-world experiments in order to achieve realistic and reliable future simulations.

Gross primary production↗

The Role of Snowmelt and Subsurface Heterogeneity in Headwater Hydrology of a Mountainous Catchment in Colorado: A Model‐Data Integration Approach

Mountainous headwater streams are sustained by both snowmelt‐driven streamflow and groundwater discharge in the Upper Colorado River Basin. However, predicting headwater stream discharge magnitude and peak flow timing is challenging in mountainous terrains, where snowmelt rates vary with vegetation type and elevation, and heterogeneous subsurface physical properties influence groundwater storage and its release. We used a model‐data integration approach to investigate the roles of snowmelt and subsurface structure in stream discharge and groundwater level. We ran an ensemble of 100 integrated surface‐subsurface hydrologic models for a mountainous headwater catchment near Crested Butte, Colorado, USA. We also evaluated and calibrated these models against observed data sets, including snow depth measurements using distributed temperature probes, stream discharge, and groundwater levels. Calibration with multiple data sources using neural density estimators has further constrained uncertainty in subsurface properties and snowmelt rates. Results indicated that observed slower snowmelt rates in evergreen forests delayed the peak flow and baseflow onset. In upstream areas with lower subsurface permeability, water was stored within the subsurface but was not released as interflow or shallow groundwater flow, and thereby not contributing to downstream streamflow during recession limb periods. Double peaks in groundwater occurred in areas with spatial subsurface heterogeneity, in our case due to the contrast between granodiorite and Mancos shale. These process‐based insights into groundwater and snowmelt dynamics in mountainous headwaters will help improve predictions of headwater hydrology.

Wang, Lijing [University of Connecticut, Storrs, C↗

Using expert systems to implement a semantic data model of a large mass storage system

The successful development of large volume data storage systems will depend not only on the ability of the designers to store data, but on the ability to manage such data once it is in the system. The hypothesis is that mass storage data management can only be implemented successfully based on highly intelligent meta data management services. There now exists a proposed mass store system standard proposed by the IEEE that addresses many of the issues related to the storage of large volumes of data, however, the model does not consider a major technical issue, namely the high level management of stored data. However, if the model were expanded to include the semantics and pragmatics of the data domain using a Semantic Data Model (SDM) concept, the result would be data that is expressive of the Intelligent Information Fusion (IIF) concept and also organized and classified in context to its use and purpose. The results are presented of a demonstration prototype SDM implemented using the expert system development tool NEXPERT OBJECT. In the prototype, a simple instance of a SDM was created to support a hypothetical application for the Earth Observing System, Data Information System (EOSDIS). The massive amounts of data that EOSDIS will manage requires the definition and design of a powerful information management system in order to support even the most basic needs of the project. The application domain is characterized by a semantic like network that represents the data content and the relationships between the data based on user views and the more generalized domain architectural view of the information world. The data in the domain are represented by objects that define classes, types and instances of the data. In addition, data properties are selectively inherited between parent and daughter relationships in the domain. Based on the SDM a simple information system design is developed from the low level data storage media, through record management and meta data management to the user interface.

Roelofs, Larry H.↗

Evolution of the ATLAS event data model for the HL-LHC

The upcoming high-luminosity run of the CERN Large Hadron Collider (HL-LHC) will yield an unprecedented volume of data. In order to process this data, the ATLAS collaboration is evolving its offline software to be able to use heterogeneous resources such as graphical processing units (GPUs) and field-programmable gate arrays (FPGAs). To reduce conversion overheads, the event data model (EDM) should be compatible with the requirements of these resources. While the ATLAS EDM has long allowed representing data as a structure of arrays, further evolution of the EDM can enable more efficient sharing of data between CPU and GPU resources. Some of this work will be summarized here, including extensions to allow controlling how memory for event data is allocated and the implementation of jagged vectors.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Data-model files associated with the manuscript titled "The Importance of Explicitly Representing the Streambed in Watershed Models" (Shuai et al., 2023 HP)

This data package contains the model inputs and outputs used in the manuscript titled "The Importance of Explicitly Representing the Streambed in Watershed Models" (Shuai et al., 2023 HP). The data.zip file contains the data used to drive the Advanced Terrestrial Simulator (ATS) model simulations. The model.zip file contains the XML input file for ATS. The notebook.zip file contains the Jupyter notebooks for pre- and post- processing model results. The figures.zip file contains the raw figures associated with the manuscript. We generated this data package in support of the manuscript and research reproducibility. The background of this study is that the streambed itself has not been explicitly represented in watershed models, although the streambed characteristics are significantly different from those of its surrounding soil. We aim to answer the following questions: 1) How do streambed properties including hydraulic conductivity, thickness and resolution impact the groundwater-surface water exchange fluxes across the streambed? 2) Is an explicit representation of streambed important in watershed modeling?; 3) Is a high-resolution streambed needed for watershed simulations?

54 ENVIRONMENTAL SCIENCES↗

An interactive environment for the analysis of large Earth observation and model data sets

Envision is an interactive environment that provides researchers in the earth sciences convenient ways to manage, browse, and visualize large observed or model data sets. Its main features are support for the netCDF and HDF file formats, an easy to use X/Motif user interface, a client-server configuration, and portability to many UNIX workstations. The Envision package also provides new ways to view and change metadata in a set of data files. It permits a scientist to conveniently and efficiently manage large data sets consisting of many data files. It also provides links to popular visualization tools so that data can be quickly browsed. Envision is a public domain package, freely available to the scientific community. Envision software (binaries and source code) and documentation can be obtained from either of these servers: ftp://vista.atmos.uiuc.edu/pub/envision/ and ftp://csrp.tamu.edu/pub/envision/. Detailed descriptions of Envision capabilities and operations can be found in the User's Guide and Reference Manuals distributed with Envision software.

Bowman, Kenneth P.↗

BIM Interoperability Tool for Improving IFC-based Model Data Exchange Between Architectural Design and Structural Analysis for Linear Piping Components

There is a growing need in the AEC industry for the digitalization of model-based data exchange in BIM workflows. However, users continue to face difficulties exchanging data between BIM-based computer-aided design (CAD) and computer-aided engineering (CAE) software, even with the open, non-proprietary data exchange format called Industry Foundation Classes (IFC). Proposed solutions in academic research focus primarily on building systems, with comparatively little attention to piping models. Therefore, this paper introduces an interoperability tool for enabling piping model data exchange between architectural and structural analysis domains. The tool is tested using a piping model created in Autodesk Revit and shows marked improvement for IFC-based CAD-to-CAE interoperability over existing practice.

97 MATHEMATICS AND COMPUTING↗

Using the NASA Giovanni DICCE Portal to Investigate Land-Ocean Linkages with Satellite and Model Data

Data-enhanced Investigations for Climate Change Education (DICCE), a NASA climate change education project, employs the NASA Giovanni data system to enable teachers to create climate-related classroom projects using selected satellite and assimilated model data. The easy-to-use DICCE Giovanni portal (DICCE-G) provides data parameters relevant to oceanic, terrestrial, and atmospheric processes. Participants will explore land-ocean linkages using the available data in the DICCE-G portal, in particular focusing on temperature, ocean biology, and precipitation variability related to El Ni?o and La Ni?a events. The demonstration includes the enhanced information for educators developed for the DICCE-G portal. The prototype DICCE Learning Environment (DICCE-LE) for classroom project development will also be demonstrated.

Acker, James G.↗