Engineering PapersSearch

SEARCH · Engineering Papers

Results for “data discover”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Electromagnetic deep-probing (100-1000 KMS) of the Earth's interior from artificial satellites: Constraints on the regional emplacement of crustal resources

Efforts continue in the development of a computer program for looking at the coupling of finite-dimensional source fields with a laterally heterogeneous Earth. An algorithm is also being developed for calculating a time-varying reference field using ground-based magnetic observatory data. It was discovered that ground-based standard magnetic observation is not as so available for the time of the MAGSAT mission as might be expected. Attempts are being made to determine the exact times and observatories from which data are avaliable.

Hermance, J. F.

Discovering Physically Meaningful Structures from Climate Extreme Data

The original proposal described an interdisciplinary team spanning UC San Diego (lead), Columbia University, and UC Irvine, with Columbia investigators including Pierre Gentine, Elias Bareinboim, and Marcus van Lier-Walqui. The proposal further specified a leadership structure in which Columbia co-investigators contributed across the three aims, with Co-PI Gentine serving as a point of contact with science teams and with responsibilities distributed across aims.

42 ENGINEERING

A Data Processing Pipeline To Extract A Knowledge Graph From Heterogeneous Data For Socio-technical Analysis Of Critical Infrastructure Influence

The code is written in Python and consists of the following pipeline that is implemented in Apache Airflow. This pipeline intends to understand the companies that are directly or indirectly involved with a type of critical infrastructure system at some point in that system's lifecycle. The pipeline takes a configuration file that specifies a list of initial companies to consider, a geographic region of interest, and a set of SEC form types as well as other data sources (e.g. CrunchBase) from which to extract entities and relations. There are four main components to this pipeline as currently implemented: Entity Extraction, Network Construction, Analysis, and Visualization. First, Entity Extraction, is implemented as the `topear-extract_organizations` Apache Airflow workflow. Given an initial query that specifies a geographic region of interest and a time interval, the software will extract CI facilities of interest and organizations that have a direct influence relationship to those facilities (e.g. ownership). During the course of the LDRD, we focused on Electric Vehicle charging stations and this information is available via the Department of Energy (DOE) database on fueling stations maintained by NREL. Within the context of the DOE CESER project, we have focused on Battery Energy Storage Systems (BESS). Second, the Network Extraction component will iteratively construct a social network graph given the set of organizations and people extracted in the previous step. Organizations (and eventually People if desired) are then fed as a query to the `topgear-construct_social_network` Apache Airflow workflow which given a set of initial companies and data sets (e.g. SEC EDGAR form types, OpenCorporates, Crunchbase). This Airflow workflow will iteratively query such data sources to discover relationships with new organizations and people. For example, this module can iteratively query SEC EDGAR for metadata that documents the number of each type of form for the given set of companies and their location. This forms metadata represents a catalog of data sources from SEC EDGAR for the extracted social network knowledge graph. The pipeline then downloads these forms from the website and saves them in a build directory for further processing. These documents are then parsed for entities and relations. Again, we note that in additional to SEC data sources, this step can also pull in information on organizations via API services such as CrunchBase and OpenCorporates or bulk data sources. At the end of this step, the resultant social network, the Critical Infrastructure network, and the edges that encode relationships between organizations and CI facilities, form the Adversarial Socio-Technical Network (ASTN) that informs the analysis. Third, the Analysis component processes these generated ASTN. Previously, that has included the ability to compare prevalence of different vendors for a given infrastructure component type across different regions as well as identify common public and private investors across those vendors. This was demonstrated for EV Charging Stations across several different metropolitan areas within an IEEE PES GridEdge publication. More recently, we have looked at ways to identify infrastructure owners and operators of BESS with the most nameplate capacity across different states as well as other indictors of risk resulting from changes in ownership over time. Finally, the Visualization component consists of an HTML/CSS/JS framework by which users can interact geospatial, operational, and organizational relationships across a given portfolio of Critical Infrastructure facilities. The objective is to provide a library of UI/UX modules that can be repurposed for stakeholder-specific dashboards. All of the modules are related via a common event model that enables UI actions in one view to percolate across the other views.

Weaver, Gabriel [Idaho National Laboratory (INL),

An Integrated and Collaborative Approach for NASA Earth Science Data

Earth science research requires coordination and collaboration across multiple disparate science domains. Data systems that support this research are often as disparate as the disciplines that they support. These distinctions can create barriers limiting access to measurements, which could otherwise enable cross-discipline Earth science. NASA's Earth Observing System Data and Information System (EOSDIS) is continuing to bridge the gap between discipline-centric data systems with a coherent and transparent system of systems that offers up to date and engaging science related content, creates an active and immersive science user experience, and encourages the use of EOSDIS earth data and services. The new Earthdata Coherent Web (ECW) project encourages cohesiveness by combining existing websites, data and services into a unified website with a common look and feel, common tools and common processes. It includes cross-linking and cross-referencing across the Earthdata site and NASA's Distributed Active Archive Centers (DAAC), and by leveraging existing EOSDIS Cyber-infrastructure and Web Service technologies to foster re-use and to reduce barriers to discovering Earth science data (http://earthdata.nasa.gov).

Murphy, K.

Active learning of ternary alloy structures and energies

Abstract Machine learning models with uncertainty quantification have recently emerged as attractive tools to accelerate the navigation of catalyst design spaces in a data-efficient manner. Here, we combine active learning with a dropout graph convolutional network (dGCN) as a surrogate model to explore the complex materials space of high-entropy alloys (HEAs). We train the dGCN on the formation energies of disordered binary alloy structures in the Pd-Pt-Sn ternary alloy system and improve predictions on ternary structures by performing reduced optimization of the formation free energy, the target property that determines HEA stability, over ensembles of ternary structures constructed based on two coordinate systems: (a) a physics-informed ternary composition space, and (b) data-driven coordinates discovered by the Diffusion Maps manifold learning scheme. Both reduced optimization techniques improve predictions of the formation free energy in the ternary alloy space with a significantly reduced number of DFT calculations compared to a high-fidelity model. The physics-based scheme converges to the target property in a manner akin to a depth-first strategy, whereas the data-driven scheme appears more akin to a breadth-first approach. Both sampling schemes, coupled with our acquisition function, successfully exploit a database of DFT-calculated binary alloy structures and energies, augmented with a relatively small number of ternary alloy calculations, to identify stable ternary HEA compositions and structures. This generalized framework can be extended to incorporate more complex bulk and surface structural motifs, and the results demonstrate that significant dimensionality reduction is possible in thermodynamic sampling problems when suitable active learning schemes are employed.

Chemistry

Assessing Videogrammetry for Static Aeroelastic Testing of a Wind-Tunnel Model

The Videogrammetric Model Deformation (VMD) technique, developed at NASA Langley Research Center, was recently used to measure displacements and local surface angle changes on a static aeroelastic wind-tunnel model. The results were assessed for consistency, accuracy and usefulness. Vertical displacement measurements and surface angular deflections (derived from vertical displacements) taken at no-wind/no-load conditions were analyzed. For accuracy assessment, angular measurements were compared to those from a highly accurate accelerometer. Shewhart's Variables Control Charts were used in the assessment of consistency and uncertainty. Some bad data points were discovered, and it is shown that the measurement results at certain targets were more consistent than at other targets. Physical explanations for this lack of consistency have not been determined. However, overall the measurements were sufficiently accurate to be very useful in monitoring wind-tunnel model aeroelastic deformation and determining flexible stability and control derivatives. After a structural model component failed during a highly loaded condition, analysis of VMD data clearly indicated progressive structural deterioration as the wind-tunnel condition where failure occurred was approached. As a result, subsequent testing successfully incorporated near- real-time monitoring of VMD data in order to ensure structural integrity. The potential for higher levels of consistency and accuracy through the use of statistical quality control practices are discussed and recommended for future applications.

Spain, Charles V.

Using Open and Interoperable Ways to Publish and Access LANCE AIRS Near-Real Time Data

The Atmospheric Infrared Sounder (AIRS) Near-Real Time (NRT) data from the Land Atmosphere Near real-time Capability for EOS (LANCE) element at the Goddard Earth Sciences Data and Information Services Center (GES DISC) provides information on the global and regional atmospheric state, with very low temporal latency, to support climate research and improve weather forecasting. An open and interoperable platform is useful to facilitate access to, and integration of, LANCE AIRS NRT data. As Web services technology has matured in recent years, a new scalable Service-Oriented Architecture (SOA) is emerging as the basic platform for distributed computing and large networks of interoperable applications. Following the provide-register-discover-consume SOA paradigm, this presentation discusses how to use open-source geospatial software components to build Web services for publishing and accessing AIRS NRT data, explore the metadata relevant to registering and discovering data and services in the catalogue systems, and implement a Web portal to facilitate users' consumption of the data and services.

Zhao, Peisheng

Crew Health and Performance Integrated Data Architecture (CHP-IDA) Project

BACKGROUND: Future Human Exploration missions introduce a new paradigm as crews move further from the resupply and near real-time ground support typical of Low Earth Orbit missions today. Without immediate support from ground-based personnel, exploration crews will be more reliant on inflight data and technology to respond to emergencies and anomalies. A data architecture to support a new generation of technologies, employing advanced analytical and predictive modeling techniques, is needed to enable crew autonomy. OVERVIEW: The Crew Health and Performance Integrated Data Architecture (CHP-IDA) project funded by NASA’s Exploration Medical Integrated Product Team (XMIPT) is laying a foundation for future in-flight informatics by providing a back-end architecture for collecting, storing, and integrating multiple sources of data generated by and around the crew. CHP-IDA provides a platform for common data models and Application Programming Interfaces to access, integrate, process, and display CHP data (e.g., environmental, exercise, medical, sleep, performance, etc.). This will facilitate the increased situation awareness and decision support required by the crew and remote support of exploration missions. This presentation will describe the currently ongoing effort to develop and evaluate a path-to-flight concept of the CHP-IDA software and its core capabilities. Current integrations will be discussed, including analytics for Extravehicular Activity metabolic rate and data ingestion from a multi-functional integrated medical device. The presentation will also provide examples of scenarios used to demonstrate the CHP-IDA through human-in-the-loop test bed activities as well as examples of appropriate system performance metrics. DISCUSSION: Today, in-flight data is often siloed, unsynchronized, and largely inaccessible in real time. Many data sets require manual entry and/or data transfer between vehicles and the ground. These issues contribute to risks in supporting exploration medical capabilities. The CHP-IDA is a back-end data system providing core capabilities needed for timely and meaningful data insights across CHP domains to crew and remote personnel to enable increased crew autonomy. Future work includes collaboration with additional CHP domains, new technology integrations, and further demonstrations of the IDA within different vehicle and communication latency contexts. LEARNING OBJECTIVES 1. The audience will understand that the CHP-IDA is a back-end system, providing a platform to facilitate access, promote decision tools, and provide meaningful insights to crew and to remote stakeholders during exploration missions. 2. The audience will gain insight into human-centered research and activities used to discover CHP domain data needs and pain points and how this information is used to guide development of the IDA.

Exploration

Induction of models under uncertainty

This paper outlines a procedure for performing induction under uncertainty. This procedure uses a probabilistic representation and uses Bayes' theorem to decide between alternative hypotheses (theories). This procedure is illustrated by a robot with no prior world experience performing induction on data it has gathered about the world. The particular inductive problem is the formation of class descriptions both for the tutored and untutored cases. The resulting class definitions are inherently probabilistic and so do not have any sharply defined membership criterion. This robot example raises some fundamental problems about induction; particularly, it is shown that inductively formed theories are not the best way to make predictions. Another difficulty is the need to provide prior probabilities for the set of possible theories. The main criterion for such priors is a pragmatic one aimed at keeping the theory structure as simple as possible, while still reflecting any structure discovered in the data.

Cheeseman, Peter

DELVE-ing into the Milky Way’s Globular Clusters: Assessing Extratidal Features in NGC 5897, NGC 7492, and Testing Detectability with Deeper Photometry

Extratidal features around globular clusters (GCs) are tracers of their disruption, stellar stream formation, and their host’s gravitational potential. However, these features remain challenging to detect due to their low surface brightness. We conduct a systematic search for such features around 19 GCs in the DECam Local Volume Exploration (DELVE) survey Data Release 2, discovering a new extra-tidal envelope around NGC 5897 and find tentative evidence for an extended envelope surrounding NGC 7492. Through a combination of dynamical modeling and analyzing synthetic stellar populations, we demonstrate these envelopes may have formed through tidal disruption. We use these models to explore the detectability of these features in the upcoming Legacy Survey of Space and Time (LSST), finding that while LSST’s deeper photometry will enhance detection significance, additional methods for foreground removal like proper motions or metallicities may be important for robust stream detection. Our results both add to the sample of globular clusters with extratidal features and provide insights on interpreting similar features in current and upcoming data.

Chiti, A. [Univ. of Chicago, IL (United States); S

Bioactivity Profiling of Chemical Mixtures for Hazard Characterization

Abstract The assessment and regulation of chemical toxicity to protect human health and the environment are done one chemical at a time and seldom at environmentally relevant concentrations. However, chemicals are found in the environment as mixtures, and their toxicity is largely unknown. Understanding the hazard posed by chemicals within the mixture is critical to enforce protective measures. Here, we demonstrate the application of bioactivity profiling of environmental water samples using the sentinel and ecotoxicology model species Daphnia to reveal the biomolecular response induced by exposure to real-world mixtures. We exposed a Daphnia strain to 30 sampled waters of the Chaobai River and measured the gene expression response profiles. Using a multiblock correlation analysis, we establish correlations between chemical mixtures identified in 30 water samples with gene expression patterns induced by these chemical mixtures. We identified 80 metabolic pathways putatively activated by mixtures of inorganic ions, heavy metals, polycyclic aromatic hydrocarbons, industrial chemicals, and a set of biocides, pesticides, and pharmacologically active substances. Our data-driven approach discovered both known bioactivity signatures with previously described modes of action and new pathways linked to undiscovered potential hazards. This study demonstrates the feasibility of reducing the complexity of real-world mixture toxicity to characterize the biomolecular effects of a defined number of chemical components based on gene expression monitoring of the sentinel species Daphnia.

Engineering

Analysis of LEAM experiment response to charged particles

The objectives of the Lunar Ejecta and Meteorites Experiment (LEAM) were to measure the long-term variations in cosmic dust influx rates and the extent and nature of the lunar ejecta. While analyzing these characteristics in the data, it was discovered that a majority of the events could not be associated with hypervelocity particle impacts of the type usually identified with cosmic dust, but could only be correlated with the lunar surface and local sun angle. The possibility that charged particles could be incident on the sensors led to an analysis of the electronics to determine if such signals could cause the large pulse height analysis (PHA) signals. A qualitative analysis of the PHA circuit showed that an alternative mode of operation existed if the input signal were composed of pulses with pulse durations very long compared to the durations for which it was designed. This alternative mode would give large PHA outputs even though the actual input amplitudes were small. This revelation led to the examination of the sensor and its response to charged particles to determine the type of signals that could be expected.

Perkins, D.

A new thermal and trajectory model for high altitude balloons

A new computer model for the prediction of the trajectory and thermal behavior of high altitude balloons has been developed. In accord with flight data, the model permits radiative emission and absorption of the lifting gas and daytime gas temperatures above that of the balloon film. It also includes ballasting, venting, and valving. Predictions obtained with the model are compared with flight data and newly discovered features are discussed.

Carlson, L. A.

New Mode For Single-Event Upsets

Report presents theory and experimental data regarding newly discovered mode for single-event upsets, (SEU's) in complementary metal-oxide/semiconductor, static random-access memories, CMOS SRAM's. SEU cross sections larger than those expected from previously known modes given rise to speculation regarding additional mode, and subsequent cross-section measurements appear to confirm speculation.

Zoutendyk, John A.

Asteroid photometry

Photoelectric light curves provide fundamental information about asteroids: rotation periods, pole orientations, shapes, and phase relations, which yield some information about the surface physical properties. This task is to carry on a program of such observations to increase the overall data base, obtain data on newly discovered asteroids, and to observe asteroids which are the subject of other complementary observations, such as occultations, radar, and infrared.

Harris, Alan W.

Autoclass: An automatic classification system

The task of inferring a set of classes and class descriptions most likely to explain a given data set can be placed on a firm theoretical foundation using Bayesian statistics. Within this framework, and using various mathematical and algorithmic approximations, the AutoClass System searches for the most probable classifications, automatically choosing the number of classes and complexity of class descriptions. A simpler version of AutoClass has been applied to many large real data sets, has discovered new independently-verified phenomena, and has been released as a robust software package. Recent extensions allow attributes to be selectively correlated within particular classes, and allow classes to inherit, or share, model parameters through a class hierarchy. The mathematical foundations of AutoClass are summarized.

Stutz, John

Bayesian classification theory

The task of inferring a set of classes and class descriptions most likely to explain a given data set can be placed on a firm theoretical foundation using Bayesian statistics. Within this framework and using various mathematical and algorithmic approximations, the AutoClass system searches for the most probable classifications, automatically choosing the number of classes and complexity of class descriptions. A simpler version of AutoClass has been applied to many large real data sets, has discovered new independently-verified phenomena, and has been released as a robust software package. Recent extensions allow attributes to be selectively correlated within particular classes, and allow classes to inherit or share model parameters though a class hierarchy. We summarize the mathematical foundations of AutoClass.

Hanson, Robin