Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “data visualizations”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 451 records · Page 25

Urban Energy Systems: Research at Oak Ridge National Laboratory

In the coming decades, our planet will witness unprecedented urban population growth in both established and emerging communities. The development and maintenance of urban infrastructures are highly energy-intensive. Urban areas are dictated by complex intersections among physical, engineered, and human dimensions that have significant implications for traffic congestion, emissions, and energy usage. In this chapter, we highlight recent research and development efforts at Oak Ridge National Laboratory (ORNL), the largest multipurpose science laboratory within the U.S. Department of Energy’s (DOE) national laboratory system, that characterizes the interactions between the human dynamics and critical infrastructures in conjunction with the integration of four distinct components: data, critical infrastructure models, and scalable computation and visualization, all within the context of physical and social systems. Discussions focus on four key topical themes: population and land use, sustainable mobility, the energy-water nexus, and urban resiliency, that are mutually aligned with DOE’s mission and ORNL’s signature science and technology capabilities. Using scalable computing, data visualization, and unique datasets from a variety of sources, the institute fosters innovative interdisciplinary research that integrates ORNL expertise in critical infrastructures including energy, water, transportation, and cyber, and their interactions with the human population.

Bhaduri, Budhu↗

TXM-Sandbox : an open-source software for transmission X-ray microscopy data analysis

A transmission X-ray microscope (TXM) can investigate morphological and chemical information of a tens to hundred micrometre-thick specimen on a length scale of tens to hundreds of nanometres. It has broad applications in material sciences and battery research. TXM data processing is composed of multiple steps. A workflow software has been developed that integrates all the tools required for general TXM data processing and visualization. The software is written in Python and has a graphic user interface in Jupyter Notebook . Users have access to the intermediate analysis results within Jupyter Notebook and have options to insert extra data processing steps in addition to those that are integrated in the software. The software seamlessly integrates ImageJ as its primary image viewer, providing rich image visualization and processing routines. As a guide for users, several TXM specific data analysis issues and examples are also presented.

36 MATERIALS SCIENCE↗

A Scoping Review of Mixed Initiative Visual Analytics in the Automation Renaissance

Artificial agents are increasingly integrated into data analysis workflows, carrying out tasks that were primarily done by humans. Our research explores how the introduction of automation recalibrates the dynamic between humans and automating technology. To explore this question, we conducted a scoping review encompassing twenty years of mixed-initiative visual analytic systems. To describe and contrast the relationship between humans and automation, we developed an integrated taxonomy to delineate the objectives of these mixed-initiative visual analytics tools, how much automation they support, and the assumed roles of humans. Here, we describe our qualitative approach of integrating existing theoretical frameworks with new codes we developed. Our analysis shows that the visualization research literature lacks consensus on the definition of mixed-initiative systems and explores a limited potential of the collaborative interaction landscape between people and automation. Our research provides a scaffold to advance the discussion of human-AI collaboration during visual data analysis. Our integrated taxonomy is available in the form of a web application on https://smonadjemi.github.io/miva.

Monadjemi, Shayan [ORNL] (ORCID:0000000293855969)↗

Open Chemistry, JupyterLab, REST, and quantum chemistry

Quantum chemistry must evolve if it wants to fully leverage the benefits of the internet age, where the worldwide web offers a vast tapestry of tools that enable users to communicate and interact with complex data at the speed and convenience of a button press. The Open Chemistry project has developed an open-source framework that offers an end-to-end solution for producing, sharing, and visualizing quantum chemical data interactively on the web using an array of modern tools and approaches. These tools build on some of the best open-source community projects such as Jupyter for interactive online notebooks, coupled with 3D accelerated visualization, state-of-the-art computational chemistry codes including NWChem and Psi4, and emerging machine learning and data mining tools such as ChemML and ANI. They offer flexible formats to import and export data, along with approaches to compare computational and experimental data.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Solar-to-Grid Public Data File for Utility-scale (UPV) and Distributed Photovoltaics (DPV) Generation, Capacity Credit, and Value

Lawrence Berkeley National Laboratory (Berkeley Lab) estimates hourly project-level generation data for utility-scale solar projects and hourly county-level generation data for residential and non-residential distributed photovoltaic (PV) systems in the seven organized wholesale markets and 10 additional Balancing Areas. To encourage its broader use, Berkeley Lab has made this data file public here at OEDI. The public project-level dataset is updated annually with data from the previous calendar year. For more information about the research project, including a technical report, briefing material, visualizations, and additional data, please visit the project homepage linked in this submission. A newer version of the data exists and can be found linked in the resources of this submission under "Solar-to-Grid Public Data File Updated 2021".

annual solar value↗

Solar-to-Grid Public Data File for Utility-scale (UPV) and Distributed Photovoltaics (DPV) Generation, Capacity Credit, and Value for 2012-2020

Lawrence Berkeley National Laboratory (Berkeley Lab) estimates hourly project-level generation data for utility-scale solar projects and hourly county-level generation data for residential and non-residential distributed photovoltaic (PV) systems in the seven organized wholesale markets and 10 additional Balancing Areas. To encourage its broader use, Berkeley Lab has made this data file public here at OEDI, covering the years 2012-2020. The public project-level dataset is updated annually with data from the previous calendar year. For more information about the research project, including a technical report, briefing material, visualizations, and additional data, please visit the project homepage linked in this submission.

annual solar value↗

Visualizing the NIOSH Pocket Guide: Open-source web application for accessing and exploring the NIOSH Pocket Guide to Chemical Hazards

The NIOSH Pocket Guide to Chemical Hazards is a trusted resource that displays key information for a collection of chemicals commonly encountered in the workplace. Entries contain chemical structures—occupational exposure limit information ranging from limits based on full-shift time-weighted averages to acute limits such as short-term exposure limits and immediately dangerous to life or health values, as well as a variety of other data such as chemical-physical properties and symptoms of exposure. The NIOSH Pocket Guide (NPG) is available as a printed, hardcopy book, a PDF version, an electronic database, and a downloadable application for mobile phones. All formats of the NIOSH Pocket Guide allow users to access the data for each chemical separately, however, the guide does not support data analytics or visualization across chemicals. This project reformatted existing data in the NPG to make it searchable and compatible with exploration and analysis using a web application. The resulting application allows users to investigate the relationships between occupational exposure limits, the range and distribution of occupational exposure limits, and the specialized sorting of chemicals by health endpoint or to summarize information of particular interest. These tasks would have previously required manual extraction of the data and analysis. The usability of this application was evaluated among industrial hygienists and researchers and while the existing application seems most relevant to researchers, the open-source code and data are amenable to modification by users to increase customization.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Data processing methods and data acquisition for samples larger than the field of view in parallel-beam tomography

Parallel-beam tomography systems at synchrotron facilities have limited field of view (FOV) determined by the available beam size and detector system coverage. Scanning the full size of samples bigger than the FOV requires various data acquisition schemes such as grid scan, 360-degree scan with offset center-of-rotation (COR), helical scan, or combinations of these schemes. Though straightforward to implement, these scanning techniques have not often been used due to the lack of software and methods to process such types of data in an easy and automated fashion. The ease of use and automation is critical at synchrotron facilities where using visual inspection in data processing steps such as image stitching, COR determination, or helical data conversion is impractical due to the large size of datasets. Here, we provide methods and their implementations in a Python package, named Algotom, for not only processing such data types but also with the highest quality possible. The efficiency and ease of use of these tools can help to extend applications of parallel-beam tomography systems.

36 MATERIALS SCIENCE↗

A Parallel Computing Infrastructure for Building Energy Simulation

In order to study grid-interactive efficient buildings, Pacific Northwest National Laboratories (PNNL) needs an infrastructure for urban-scale building energy modeling. Such an infrastructure should be fast, scalable, and easy-to-use. Given a set of data from the Energy Information Administration’s Commercial Building Energy Consumption Survey (CBECS) and tool to translate survey data into simulation inputs, this project aimed to conduct the simulation of the entire dataset in parallel. Before running the simulations, the necessary software was bundled into a container for use on the PNNL supercomputing network. Then, the parallel simulation workflow was designed using GNU Make, a file creation software, and submitted to a supercomputing partition which could run hundreds of simulations simultaneously. The EnergyPlus simulations output hourly electric meter data for each CBECS sample, which represents the electricity consumption of similar commercial buildings across the United States. Analyzing and visualizing the meter data is important to the future of the work, and this project wrote code to make common analysis methods simple, fast, and accessible. Moving forwards, the model will need to be expanded to include data from other sources and its accuracy will need to be improved and eventually validated.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

House Advantage or House of Cards? Stacking the Deck for Data Videos Leads to Null Results: Preprint

Videos are becoming a ubiquitous means of sharing information on social media platforms. In response, data videos - short clips combining visualization with dynamic storytelling, audio descriptions, and spatial referencing - have gained popularity for communicating data. These affordances suggest that data videos might communicate data patterns, trends, and concepts more effectively than static visualizations, enhancing comprehension. However, existing research has not systematically tested this claim. To address this gap, we conducted three controlled studies to measure comprehension differences between data videos and static visualizations. Despite leveraging visual cues and audio explanations, no data video led to significantly better comprehension than an analogous static visualization. Our results suggest data videos are not categorically better and that future research should examine the tradeoffs between their engagement benefits and costs.

97 MATHEMATICS AND COMPUTING↗

A model-independent data assimilation (MIDA) module and its applications in ecology

Abstract. Models are an important tool to predict Earth system dynamics. An accurate prediction of future states of ecosystems depends on not only model structures but also parameterizations. Model parameters can be constrained by data assimilation. However, applications of data assimilation to ecology are restricted by highly technical requirements such as model-dependent coding. To alleviate this technical burden, we developed a model-independent data assimilation (MIDA) module. MIDA works in three steps including data preparation, execution of data assimilation, and visualization. The first step prepares prior ranges of parameter values, a defined number of iterations, and directory paths to access files of observations and models. The execution step calibrates parameter values to best fit the observations and estimates the parameter posterior distributions. The final step automatically visualizes the calibration performance and posterior distributions. MIDA is model independent, and modelers can use MIDA for an accurate and efficient data assimilation in a simple and interactive way without modification of their original models. We applied MIDA to four types of ecological models: the data assimilation linked ecosystem carbon (DALEC) model, a surrogate-based energy exascale earth system model: the land component (ELM), nine phenological models and a stand-alone biome ecological strategy simulator (BiomeE). The applications indicate that MIDA can effectively solve data assimilation problems for different ecological models. Additionally, the easy implementation and model-independent feature of MIDA breaks the technical barrier of applications of data–model fusion in ecology. MIDA facilitates the assimilation of various observations into models for uncertainty reduction in ecological modeling and forecasting.

58 GEOSCIENCES↗

Guiding the choice of informatics software and tools for lipidomics research applications

Progress in mass spectrometry lipidomics has led to a rapid proliferation of studies across biology and biomedicine. These generate extremely large raw datasets requiring sophisticated solutions to support automated data processing. To address this, numerous software tools have been developed and tailored for specific tasks. However, for researchers, deciding which approach best suits their application relies on ad hoc testing, which is inefficient and time consuming. Here we first review the data processing pipeline, summarizing the scope of available tools. Next, to support researchers, LIPID MAPS provides an interactive online portal listing open-access tools with a graphical user interface. This guides users towards appropriate solutions within major areas in data processing, including (1) lipid-oriented databases, (2) mass spectrometry data repositories, (3) analysis of targeted lipidomics datasets, (4) lipid identification and (5) quantification from untargeted lipidomics datasets, (6) statistical analysis and visualization, and (7) data integration solutions. Detailed descriptions of functions and requirements are provided to guide customized data analysis workflows.

59 BASIC BIOLOGICAL SCIENCES↗

Development of Saturated Zone Three-Dimensional Initial Condition Plumes for the Composite Analysis and Cumulative Impacts Evaluation Modeling

The purpose of this environmental calculation file (ECF) is to document and present the methodology, input data, and results of extending two-dimensional (2-D) plume depictions documented in DOE/RL- 2017-66, Hanford Site Groundwater Monitoring for 2017, into three dimensions (3-D) for select contaminants of interest (COIs) in the 200-BP-5, 200-UP-1, 200-ZP-1, and 200-PO-1 Groundwater Operable Units at the Hanford Site. Plumes with sufficient available vertical profile data were interpolated in 3-D using the Leapfrog Geo®1 geologic modeling platform, which includes an interpolation engine for analyzing and visualizing 3-D data. Results of these interpolations are intended for use in developing inputs for transport modeling to support forecasts of Central Plateau contaminant plume fate and transport.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

Synthetic Streamflow Datasets to Support Emulation of Water Allocations via LSTM

This archive is the data companion to the bonney_et-al_2026_erc metarepo which generates synthetic data, trains an LSTM model, and generates performance metrics on the trained model. While the generation of the synthetic data is fully reprodicible, it is a computationally expensive process. This data archive contains the synthetic datasets needed for training and testing an LSTM model and reproduction of figures and tables. In addition, supplemenatary data products generating and visualizing results is also included, such as geospatial data for the basin. Contents There are two high level directories: `WRAP_archive/` and `repo_data/`. The `WRAP_archive` directory contains compressed intermediate dataproducts from the dataset generation workflow (marked as "I_Dataset_Generation" in the metarepo). These data products are not required by any scripts in the metarepo, but they are archived as they are expensive to generate and may have useful information for other analyses. The `repo_data` directory contains the necessary data for reproducing the workflow in the metarepo and should be decompressed and moved into the top level of the metarepo. Additional details are provided in README.md.

drought↗

XtalCAMP: a comprehensive program for the analysis and visualization of scanning Laue X-ray micro-/nanodiffraction data

XtalCAMP is a software package based on the MATLAB platform, which is suitable for, but not limited to, the analysis and visualization of scanning Laue X-ray micro-/nanodiffraction data. The main objective of the software is to provide complementary functionalities to the Laue indexing software packages used at several synchrotron beamlines. Here, the graphical user interfaces allow the easy analysis of characteristic microstructure features, including real-time intensity mapping for a quick examination of phase, grain and defect distribution, 2D color-coded mapping of microstructural properties from the output of other Laue indexing software, crystal orientation visualization, grain boundary characterization based on orientation/misorientation calculation, principal strain/stress analysis, and strain ellipsoid representation, as well as a series of additional toolkits. As an example, XtalCAMP is applied to the microstructural investigation of a solution-heat-treated Ni-based superalloy manufactured using a laser 3D-printing technique, and a deformed natural quartzite from Val Bregaglia in the Central Alps.

36 MATERIALS SCIENCE↗

Marmot

Marmot is a data formatting and visualization tool for production cost modelling results. It provides an efficient way to view PLEXOS production cost modeling results quickly, while also creating publication ready figures and data tables.

Levie, Daniel↗

ICAT: The Interactive Corpus Analysis Tool

The Interactive Corpus Analysis Tool (ICAT) is a Python library for creating dashboards to explore textual datasets and build simple binary classification models to help filter through them and focus on entries of interest. This tool uses a form of interactive machine learning (IML), a paradigm of “machine teaching” (Simard et al., 2017) that sits at the intersection of the fields of human computer interaction (HCI), visual analytics, and machine learning. The intent of ICAT is to allow subject matter experts (SME) with limited to no experience in machine learning to benefit from an iterative human-in-the-loop (HITL) approach to building their own model without needing to understand the details of the underlying algorithm. This interactivity is achieved by allowing the user to create features, label data points, and visually manipulate a representation of the features to manually cluster and investigate data, while a model is trained on the fly based on these actions. ICAT is built on top of the Panel (Holoviz, 2018) library, using a combination of Vega, a custom IPyWidget using D3, and ipyvuetify, and is intended to be used inside of a Jupyter environment.

Martindale, Nathan [Oak Ridge National Laboratory ↗

A primer on forest structure measurement with lidar for ecologists

Light detection and ranging (lidar) technology has fundamentally advanced the way we measure forest structure, facilitating new insights into ecological processes. Lidar for forest ecology applications is deployed on multiple types of platforms that operate from the ground, air, or space, and each has associated strengths and limitations. Ideally, the choice of what kind of lidar to use in a particular study should be guided by the ecological question of interest; however, practical considerations of cost, data availability, and processing tools can be equally important. This synthesis is a practical introduction to how different lidar platforms characterize forest structure (e.g., tree size/location, wood volume, branching structure, aboveground biomass, leaf properties), designed for a general audience of ecologists (not remote sensing scientists) seeking an accessible introduction to the use of lidar. We also provide examples of novel ecological insights from recent lidar research and describe current limitations and areas of expected improvement. Last, we include an appendix of data collected from terrestrial, mobile, unoccupied aerial system, airplane, and satellite lidar platforms within a common temperate forest area, with associated code to allow new lidar users to visualize and manipulate data in R.

Cushman, KC [ORNL] (ORCID:0000000234641151)↗