Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “common information model”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

GriddingMachine, a database and software for Earth system modeling at global and regional scales

Land and Earth system modeling is moving towards more explicit biophysical representations, requiring increasing variety of datasets for initialization and benchmarking. However, researchers often have difficulties in identifying and integrating non-standardized datasets from various sources. We aim towards a standardized database and one-stop distribution method of global datasets. Here, we present the GriddingMachine as (1) a database of global-scale datasets commonly used to parameterize or benchmark the models, from plant traits to vegetation indices and geophysical information and (2) a cross-platform open source software to download and request a subset of datasets with only a few lines of code. The GriddingMachine datasets can be accessed either manually through traditional HTTP, or automatically using modern programming languages including Julia, Matlab, Octave, Python, and R. The GriddingMachine collections can be used for any land and Earth modeling framework and ecological research at the regional and global scales, and the number of datasets will continue to grow to meet the increasing needs of research communities.

58 GEOSCIENCES↗

Analytical Sensitivity Analysis of a Spent Nuclear Fuel Cask

Here, we report nuclear science and engineering is a field increasingly dominated by computational studies resulting from increasingly powerful computational tools. As a result, analytical studies, which previously pioneered nuclear engineering, are increasingly viewed as secondary or unnecessary. However, analytical solutions to reduced-fidelity models can provide important information concerning the underlying physics of a problem and aid in guiding computational studies. Similarly, there is increased interest in sensitivity analysis studies. These studies commonly use computational tools. However, providing a complementary sensitivity study of relevant analytical models can lead to a deeper analysis of a problem. This work provides the analytical sensitivity analysis of the one-dimensional (1D) cylindrical mono-energetic neutron diffusion equation using the forward sensitivity analysis procedure (FSAP) developed by Cacuci. Further, these results are applied to a reduced-fidelity model of a spent nuclear fuel cask, demonstrating how computational analysis might be improved with a complementary analytic sensitivity analysis.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

Extraction and Analysis of Time Series Data from Building Automation Systems Using Large Language Models

Semantic schemas like Haystack 4, Brick and ASHRAE standard 223 enable the structured, standardized, and machine-readable representation of building data, facilitating interoperability, data integration, and advanced analytics. However, extracting information from these models requires specialized expertise in SPARQL and other programming languages, skills that are not commonly found among building professionals. Recent advancements in Large Language Models (LLMs), such as ChatGPT, enable the construction of queries using natural language, making it easier for individuals to interact with these systems in a manner that resembles everyday speech. However, these methods have not yet been tested on building semantic ontologies. This paper introduces a novel workflow and tool for enabling users to ask questions about a specific building's data, using natural language and receive answers automatically generated by GPT-4o. Our approach integrates semantic ontologies with advanced LLM capabilities to automate three critical steps: (1) generating SPARQL queries to retrieve time series references from ontological models, (2) extracting the corresponding time series data from the Building Automation System, and (3) performing computations and visualizations tailored to the user's query. The proposed method simplifies access to BAS data, allowing both domain experts and non-specialists to conduct sophisticated analyses without needing extensive technical knowledge of semantic web technologies. By demonstrating this pipeline, we facilitate more accessible and scalable data-driven decision-making in building operations and management.

Mulayim, Ozan Baris↗

Extraction and Analysis of Time Series Data from Building Automation Systems Using Large Language Models

Semantic schemas like Haystack 4, Brick and ASHRAE standard 223 enable the structured, standardized, and machine-readable representation of building data, facilitating interoperability, data integration, and advanced analytics. However, extracting information from these models requires specialized expertise in SPARQL and other programming languages, skills that are not commonly found among building professionals. Recent advancements in Large Language Models (LLMs), such as ChatGPT, enable the construction of queries using natural language, making it easier for individuals to interact with these systems in a manner that resembles everyday speech. However, these methods have not yet been tested on building semantic ontologies. This paper introduces a novel workflow and tool for enabling users to ask questions about a specific building's data, using natural language and receive answers automatically generated by GPT-4o. Our approach integrates semantic ontologies with advanced LLM capabilities to automate three critical steps: (1) generating SPARQL queries to retrieve time series references from ontological models, (2) extracting the corresponding time series data from the Building Automation System, and (3) performing computations and visualizations tailored to the user's query. The proposed method simplifies access to BAS data, allowing both domain experts and non-specialists to conduct sophisticated analyses without needing extensive technical knowledge of semantic web technologies. By demonstrating this pipeline, we facilitate more accessible and scalable data-driven decision-making in building operations and management.

Mulayim, Ozan Baris↗

Technical note: AQMEII4 Activity 1: evaluation of wet and dry deposition schemes as an integral part of regional-scale air quality models

We present in this technical note the research protocol for phase 4 of the Air Quality Model Evaluation International Initiative (AQMEII4). This research initiative is divided into two activities, collectively having three goals: (i) to define the current state of the science with respect to representations of wet and especially dry deposition in regional models, (ii) to quantify the extent to which different dry deposition parameterizations influence retrospective air pollutant concentration and flux predictions, and (iii) to identify, through the use of a common set of detailed diagnostics, sensitivity simulations, model evaluation, and reduction of input uncertainty, the specific causes for the current range of these predictions. Activity 1 is dedicated to the diagnostic evaluation of wet and dry deposition processes in regional air quality models (described in this paper), and Activity 2 to the evaluation of dry deposition point models against ozone flux measurements at multiple towers with multiyear observations (to be described in future submissions as part of the special issue on AQMEII4). The scope of this paper is to present the scientific protocols for Activity 1, as well as to summarize the technical information associated with the different dry deposition approaches used by the participating research groups of AQMEII4. In addition to describing all common aspects and data used for this multi-model evaluation activity, most importantly, we present the strategy devised to allow a common process-level comparison of dry deposition obtained from models using sometimes very different dry deposition schemes. The strategy is based on adding detailed diagnostics to the algorithms used in the dry deposition modules of existing regional air quality models, in particular archiving diagnostics specific to land use–land cover (LULC) and creating standardized LULC categories to facilitate cross-comparison of LULC-specific dry deposition parameters and processes, as well as archiving effective conductance and effective flux as means for comparing the relative influence of different pathways towards the net or total dry deposition. This new approach, along with an analysis of precipitation and wet deposition fields, will provide an unprecedented process-oriented comparison of deposition in regional air quality models. Examples of how specific dry deposition schemes used in participating models have been reduced to the common set of comparable diagnostics defined for AQMEII4 are also presented.

54 ENVIRONMENTAL SCIENCES↗

Three-Receiver Quantum Broadcast Channels: Classical Communication with Quantum Non-unique Decoding

In network communication, it is common in broadcasting scenarios for there to exist a hierarchy among receivers based on information they decode due, for example, to different physical conditions or premium subscriptions. This hierarchy may result in varied information quality, such as higher-quality video for certain receivers. This is modeled mathematically as a degraded message set, indicating a hierarchy between messages to be decoded by different receivers, where the default quality corresponds to a common message intended for all receivers, a higher quality is represented by a message for a smaller subset of receivers, and so forth. We extend these considerations to quantum communication, exploring three-receiver quantum broadcast channels with two- and three-degraded message sets. Our technical tool involves employing quantum non-unique decoding, a technique we develop by utilizing the simultaneous pinching method. Here, we construct one-shot codes for various scenarios and find achievable rate regions relying on various quantum Rényi mutual information error exponents. Our investigation includes a comprehensive study of pinching across tensor product spaces, presenting our findings as the asymptotic counterpart to our one-shot codes. By employing the non-unique decoding, we also establish a simpler proof to Marton’s inner bound for two-receiver quantum broadcast channels without the need for more involved techniques. Additionally, we derive no-go results and demonstrate their tightness in special cases.

Salek, Farzin [Technical University of Munich (Ger↗

Developing new pathways for energy and environmental decision-making in India: a review

Abstract India faces a dual challenge of economic development and responding to climate change. Although India’s per capita emissions are well below global average, the country is one of the world’s largest greenhouse gas emitters. Indian policymakers and stakeholders require high-quality data and research to assess low-emissions, sustainable development strategies. Peer-reviewed literature is a key source of this information and also a key venue for conversation amongst research leaders. This paper examines the recent peer-reviewed literature on India’s 2030 and 2050 pathways. We conducted a systematic literature review to identify key quantitative national modeling studies. From the 34 studies identified, we synthesized scenario data to draw common conclusions and identify critical research gaps. The main focus was on examining the coverage and the state of information available on low-carbon pathways. Overall, we find a few scenarios that are potentially consistent with a 2070 net-zero goal, but more limited assessment of pathways to reach net-zero emissions before this date. Mitigation pathways with greater ambition are required across all energy sectors to ensure a smooth transition to net-zero emissions by or before 2070. The scenarios confirm that reducing emissions to below 2 GtCO 2 yr −1 by mid-century would necessitate significant transformations of the Indian energy sector, such as, a decrease in unabated coal power capacity, transportation modal shift, and industrial process switching. The assessment also finds substantial differences in final energy estimates reported across studies, particularly in transportation. The lack of consistency in, and transparency about underlying drivers, assumptions, and even outputs across studies points to the critical need for the sorts of coordinated, multi-model studies that have proven exceptionally valuable for decision makers in other major emitting countries.

54 ENVIRONMENTAL SCIENCES↗

Constraining Physical Models at Gigabar Pressures

High-energy-density (HED) experiments in convergent geometry are able to test physical models at pressures beyond hundreds of millions of atmospheres. The measurements from these experiments are generally highly integrated and require unique analysis techniques to procure quantitative information. This work describes a methodology to constrain the physics in convergent HED experiments by adapting the methods common to many other fields of physics. As an example, a mechanical model of an imploding shell is constrained by data from a thin-shelled direct-drive exploding-pusher experiment on the OMEGA Laser System using Bayesian inference, resulting in the reconstruction of the shell dynamics and energy transfer during the implosion. The model is tested by analyzing synthetic data from a 1-D hydrodynamics code and is sampled using a Markov chain Monte Carlo to generate the posterior distributions of the model parameters. The goal of this work is to demonstrate a general methodology that can be used to draw conclusions from a wide variety of HED experiments.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Leveraging Open-Source Tools for Collaborative Macro-energy System Modeling Efforts

The authors are founding team members of a new effort to develop an Open Energy Outlook for the United States. The effort aims to apply best practices of policy-focused energy system modeling, ensure transparency, build a networked community, and work toward a common purpose: examining possible US energy system futures to inform energy and climate policy efforts. Individual author biographies can be found on the project website: https://openenergyoutlook.org/.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Evaluation of the Self Retrieval Augmented Generation Technique on Common Security Advisory Framework Data

This small experimental report evaluates a variation of Retrieval Augmented Generation (RAG), called Self-RAG. This method uses a generative language model that incorporates retrieved facts into its generation and is explicitly trained to be able to determine whether retrieved information is enough to answer the input query, with a user-defined threshold for confidence. We performed an experiment using data from the publicly available CISA Common Security Advisory Framework (CSAF) repository (https://github.com/cisagov/CSAF) as the database of facts to be used in retrieval. Qualitative results from the experiment demonstrate that the Self-RAG method has some ability to provide reasonable answers to queries that are in the dataset and will often ignore irrelevant information when asked outside of domain questions (e.g., general facts). In settings with deliberately confusing questions (the question is within domain, but asks about a fabricated advisory), it was able to refuse 40% of the time without further adjustments to the original framework. While this performance is not sufficient for current practical use, further improvements to data formatting, disambiguating results, and leveraging threshold values could improve performance significantly. However, evaluating this will require more extensive evaluations on larger datasets and potentially better models.

97 MATHEMATICS AND COMPUTING↗

FolpsD: combining EFT and phenomenological approaches for joint power spectrum and bispectrum analyses

We present a theoretical model for the power spectrum and bispectrum of galaxy clustering that exploits the complementarity between small-scale power spectrum information and large-scale bispectrum measurements. We extend the FOLPS code by combining its one-loop EFT galaxy power spectrum with a tree-level galaxy bispectrum projected onto the tripolar spherical harmonics (Sugiyama) basis. To access additional small-scale information, we also consider a line-of-sight damping factor in both statistics, mirroring approaches commonly used in studies of redshift-space distortions. We test the model using DESI DR2 galaxy mocks. Even without damping, the joint analysis of the EFT power spectrum and bispectrum significantly improves constraints and reduces parameter degeneracies relative to power spectrum analyses alone. For LRG-like samples, including the damping further extends the range beyond $k\sim 0.3 \,h \text{Mpc}^{-1}$ in the power spectrum and $k \sim 0.24 \,h \text{Mpc}^{-1}$ in the bispectrum without introducing statistically significant parameter biases. This leads to up to $\sim 30\%$ tighter constraints on $A_s$ and $ω_{cdm}$. For low signal-to-noise tracers such as QSOs, however, the damping parameters are weakly constrained and can absorb noise fluctuations, leading to shifts in inferred parameters. Similar limitations may arise in models where cosmological information is encoded in power-spectrum shape features degenerate with the damping, such as scenarios with massive neutrinos. In contrast, for $w_0w_a$CDM we obtain $15\%$ and $21\%$ tighter constraints on $w_0$ and $w_a$, respectively, yielding a deviation from constant dark energy at slightly more than the $1σ$ level using full-shape information alone. The code is publicly available at https://github.com/cosmodesi/FolpsD

Bansal, P. [Michigan U., MCTP; Michigan U.] (ORCID↗

Maven: a multimodal foundation model for supernova science

Abstract A common setting in astronomy is the availability of a small number of high-quality observations, and larger amounts of either lower-quality observations or synthetic data from simplified models. Time-domain astrophysics is a canonical example of this imbalance, with the number of supernovae observed photometrically outpacing the number observed spectroscopically by multiple orders of magnitude. At the same time, no data-driven models exist to understand these photometric and spectroscopic observables in a common context. Contrastive learning objectives, which have grown in popularity for aligning distinct data modalities in a shared embedding space, provide a potential solution to extract information from these modalities. We present Maven, the first foundation model for supernova science. To construct Maven, we first pre-train our model to align photometry and spectroscopy from 0.5 M synthetic supernovae using a contrastive objective. We then fine-tune the model on 4702 observed supernovae from the Zwicky transient facility. Maven reaches state-of-the-art performance on both classification and redshift estimation, despite the embeddings not being explicitly optimized for these tasks. Through ablation studies, we show that pre-training with synthetic data improves overall performance. In the upcoming era of the Vera C. Rubin observatory, Maven will serve as a valuable tool for leveraging large, unlabeled and multimodal time-domain datasets.

Zhang, Gemma (ORCID:0000000280198082)↗

Bayesian reduced-order deep learning surrogate model for dynamic systems described by partial differential equations

We propose a reduced-order deep-learning surrogate model for dynamic systems described by time-dependent partial differential equations. This method employs space–time Karhunen–Loève expansions (KLEs) of the state variables and space-dependent KLEs of space-varying parameters to identify the reduced (latent) dimensions. Subsequently, a deep neural network (DNN) is used to map the parameter latent space to the state variable latent space. An approximate Bayesian method is developed for uncertainty quantification (UQ) in the proposed KL-DNN surrogate model. The KL-DNN method is tested for the linear advection–diffusion and nonlinear diffusion equations, and the Bayesian approach for UQ is compared with the deep ensembling (DE) approach, commonly used for quantifying uncertainty in DNN models. It was found that the approximate Bayesian method provides a more informative distribution of the PDE solutions in terms of the coverage of the reference PDE solutions (the percentage of nodes where the reference solution is within the confidence interval predicted by the UQ methods) and log predictive probability. The DE method is found to underestimate uncertainty and introduce bias. For the nonlinear diffusion equation, we compare the KL-DNN method with the Fourier Neural Operator (FNO) method and find that KL-DNN is 10% more accurate and needs less training time than the FNO method.

97 MATHEMATICS AND COMPUTING↗

Workshop on Addressing Rigor and Reproducibility in Thermal, Heterogeneous Catalysis

Heterogeneous catalysis has long served as the bedrock of the manufacturing of energy carriers, fuels and chemicals, and various technologies for pollution abatement. The significant complexity and variability spanning the entire breadth of catalyst material properties, synthesis methods, characterization techniques, and evaluation procedures, has focused attention on the need to establish community-accepted best practices for ensuring high-quality, benchmarked, and reproducible data. In addition, increased societal urgency to transition to clean energy and reduce greenhouse gas concentrations has incentivized interdisciplinary, convergent, and translational approaches to catalysis research in recent years. Research engineers and scientists with expertise cutting broadly across materials science, chemical synthesis, interfacial science, spectroscopy, and methods of data science and computational simulation, all bring diverse and important perspectives to catalysis research, but often with little awareness of the complexity of catalytic systems, especially in their working environment. As has already occurred in other scientific fields, there has been growing recognition and consensus in the heterogeneous catalysis research community that mechanisms are needed to improve the rigor and reproducibility (R&R) of experimental measurements, to ensure alignment of the broader research community with a common core of best practices specific to the realization of high-quality catalysis research. Similarly, the field is moving rapidly toward computationally informed and data science-driven catalyst design, but the success of implementing such predictive tools hinges on model training and validation rooted in rigorously obtained and reproducible experimental data that are benchmarked to common specifications. As such, this workshop was convened to prepare a report summarizing best practices for reporting data and performing experiments that researchers can use to benchmark, validate, and reproduce data in specific sub-fields of thermal, heterogeneous catalysis. Additionally, we discussed recommendations for future actions that may improve R&R in this field. The workshop organizers and participants include a diverse range of catalysis researchers from various employment sectors (e.g., academia, industry, national laboratory), institutional mission and resources (e.g., PhD-granting research universities, non-PhD-granting teaching universities), career stage (e.g., early, mid and late-career), technical expertise, and demographic background. This diverse group was involved in the discussion of workshop agenda items, writing this report, and discussing possible future action items for the community to consider, which helped ensure that a broad range of perspectives were captured in the description of the problems at hand and the creation of actionable solutions that may be effectively adopted by the diverse practitioners in catalysis research. Importantly, this group of workshop participants also included very early career researchers (e.g., senior PhD students, postdoctoral scholars) who will become the next generation of scientific leaders in various sectors, thus capturing emerging perspectives of newcomers to the field to shape its future while positively impacting the development of its future workforce. We envision that this effort will help advance the field of catalysis science by improving the rigor and reproducibility of experimental data collected by current researchers and future newcomers to the field, which is of broad importance to health and vitality of any scientific discipline. Therefore, best practices identified in this endeavor for thermal heterogeneous catalysis can be translated to such efforts in other areas of catalysis and other scientific fields involving the study of materials, and vice versa. We also envision this to be an ongoing effort, with future workshops that are convened to discuss issues of rigor and reproducibility on technical topics that were unable to be covered in this workshop due to its scope limitations, and as emerging methods and materials become more prevalent in the research community.

36 MATERIALS SCIENCE↗

Generation of Data-Driven Expected Energy Models for Photovoltaic Systems

Although unique expected energy models can be generated for a given photovoltaic (PV) site, a standardized model is also needed to facilitate performance comparisons across fleets. Current standardized expected energy models for PV work well with sparse data, but they have demonstrated significant over-estimations, which impacts accurate diagnoses of field operations and maintenance issues. This research addresses this issue by using machine learning to develop a data-driven expected energy model that can more accurately generate inferences for energy production of PV systems. Irradiance and system capacity information was used from 172 sites across the United States to train a series of models using Lasso linear regression. The trained models generally perform better than the commonly used expected energy model from international standard (IEC 61724-1), with the two highest performing models ranging in model complexity from a third-order polynomial with 10 parameters (Radj2 = 0.994) to a simpler, second-order polynomial with 4 parameters (Radj2=0.993), the latter of which is subject to further evaluation. Subsequently, the trained models provide a more robust basis for identifying potential energy anomalies for operations and maintenance activities as well as informing planning-related financial assessments. We conclude with directions for future research, such as using splines to improve model continuity and better capture systems with low (≤1000 kW DC) capacity.

14 SOLAR ENERGY↗

Improved information criteria for Bayesian model averaging in lattice field theory

Bayesian model averaging is a practical method for dealing with uncertainty due to model specification. Use of this technique requires the estimation of model probability weights. Here, we revisit the derivation of estimators for these model weights. Use of the Kullback-Leibler divergence as a starting point leads naturally to a number of alternative information criteria suitable for Bayesian model weight estimation. We explore three such criteria, known to the statistics literature before, in detail: a Bayesian analog of the Akaike information criterion which we call the BAIC, the Bayesian predictive information criterion, and the posterior predictive information criterion (PPIC). We compare the use of these information criteria in numerical analysis problems common in lattice field theory calculations. We find that the PPIC has the most appealing theoretical properties and can give the best performance in terms of model-averaging uncertainty, particularly in the presence of noisy data, while the BAIC is a simple and reliable alternative.

97 MATHEMATICS AND COMPUTING↗

Developing an Energy Service Interface Specification

Developing an Energy Service Interface (ESI) specification requires engaging a community of stakeholders including grid operators, Information and Communication Technology implementors, integrators, and finally standards bodies who will define an interface that respects and boundaries of ownership and roles of responsibility in order to activate millions of Distributed Energy Resource for the provision of grid services. By applying Interoperability Maturity Model Criteria and ESI principles to common grid-DER service use cases, the Grid Modernization Lab Consortium team will engage subject matter experts to develop a specification, with an eventual goal of informing development of ESI compliant profiles or standards. The ESI Specification is intended to specify the characteristics, attributes, or qualities that need to be addressed in ESI compliant standards or profiles. This includes addressing interoperability criteria and the service-performance style of the interface.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

The "PVLib" of Degradation: PVDeg

The Photovoltaic (PV) industry constantly aims for lower costs through higher-efficiency cells, improved module designs, and improvements in durability. This leads to the use of new materials, designs, and manufacturing processes, and not always with a sufficient amount of durability testing. To help drive down costs there is a desire to create modules that will last for up to 50 years of service life. To accomplish this, every degradation mode and mechanism must be identified and either eliminated or otherwise mitigated. This involves the extrapolation of laboratory results to the field conditions. There is a need to organize the existing degradation data into an accessible format and to provide industry relevant tools for extrapolation from laboratory to field conditions. While the basic equations used to model degradation are sometimes very simple, the full analysis involves calculations are cumbersome but ubiquitous for many degradation processes. A simplified, modeling framework to accomplish these repetitive processes will facilitate the analysis to help researchers keep up with the rapid pace of technological changes. In this talk, we will describe our progress creating the open-source tool PVDeg. This tool can be used to search for and analyze degradation information and extrapolate PV module performance and durability to field exposure. PVDeg simplifies many of the common foundational computational operations for obtaining meteorological data and using it to generate a model of the PV deployment. This prediction tool repository also contains various degradation models as well as a library of material parameters suitable for estimating the durability assessment of materials and components. We use an integration pipeline approach that allows us to leverage weather data from the National Solar Radiation Database, and other weather sources, to perform geospatial degradation analysis in the US and worldwide. We hope to become a repository that can be used for weathering and degradation analysis for various applications beyond the PV industry. During the talk, we will provide the PVPMC attendees the opportunity to interact with the tool via a Google Collab tutorial they can run on their phones or laptops.

durability↗