Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “databases”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 379 records · Page 21

ORNL National Hydropower Fish Passage Database

Fish passage facilities are used to mitigate impacts of hydropower dams to migratory fish in rivers, but information on the location, types, and characteristics of this infrastructure is incomplete at a national scale. Researchers at Oak Ridge National Laboratory (ORNL) partnered with fish passage engineers and hydropower experts from the US Fish and Wildlife Service (USFWS), the National Oceanic Atmospheric Administration’s National Marine Fisheries Service (NOAA NMFS), and the Low Impact Hydropower Institute (LIHI) to create the first national scale database of fish passage infrastructure at US hydropower developments. This database consists of the ORNL_Fish_Passage_Dataset.zip file with 13 individual .csv files that contain information on fish passage facility engineering characteristics, targeted fish species, operational schedule, and costs. which is of great value to a diverse range of stakeholders. This information was collected between December 2023 and July 2025 from project partners, other hydropower stakeholders, online datasets available for download, and a stakeholder questionnaire Information. This data resource addresses a large gap in knowledge of the deployment of fish passage technology and is freely available to members of the hydropower community, including federal and state regulators and resource agencies, non-governmental organizations (NGOs), industry, and other user groups to support project planning and regulatory (re)licensing activities. This database supports the US Department of Energy Water Power Technologies Office objective to develop decision support tools and data resources that improve environmental performance and ensure hydropower’s long-term value to the American public.

Matson, Paul [Oak Ridge National Laboratory (ORNL)↗

An experimental database of cell performance for vanadium redox flow battery

The continual growth in energy demand has resulted in the deployment of renewable energy generators to reduce the impact of fossil fuel dependence. However, these generators often suffer from intermittency and require energy storage when there is over-generation and the subsequent release of this stored energy at high demand. One promising energy storage technology which can provide a solution to improve energy management and grid stability, is the redox flow battery. Among the numerous flow battery systems, vanadium redox flow battery is the most iconic solution to large scale energy storage, giving a more efficient link between energy production, especially from renewables, and energy demand. The aim of the current database is to characterize the performance of the cell design and to provide training/validation data for physical model or data-driven model. The database includes hundreds of experimental cell performance data of vanadium redox flow battery with various current densities for multiple charge-discharge cycles. All the cell parameters, chemical parameters, material parameters, operation parameters and thermodynamic parameters of the cell system are listed. Coulomb, voltaic and energy efficiencies are also provided. The database will be helpful for researchers in the field of redox flow batteries.

Gao, Peiyuan↗

A Comprehensive Urine Proteome Database Generated From Patients With Various Renal Conditions and Prostate Cancer

Urine proteins can serve as viable biomarkers for diagnosing and monitoring various diseases. A comprehensive urine proteome database, generated from a variety of urine samples with different disease conditions, can serve as a reference resource for facilitating discovery of potential urine protein biomarkers. Herein, we present a urine proteome database generated from multiple datasets using 2D LC-MS/MS proteome profiling of urine samples from healthy individuals (HI), renal transplant patients with acute rejection (AR) and stable graft (STA), patients with non-specific proteinuria (NS), and patients with prostate cancer (PC). A total of ~28,000 unique peptides spanning ~2,200 unique proteins were identified with a false discovery rate of <0.5% at the protein level. Over one third of the annotated proteins were plasma membrane proteins and another one third were extracellular proteins according to gene ontology analysis. Ingenuity Pathway Analysis of these proteins revealed 349 potential biomarkers in the literature-curated database. Forty-three percentage of all known cluster of differentiation (CD) proteins were identified in the various human urine samples. Interestingly, following comparisons with five recently published urine proteome profiling studies, which applied similar approaches, there are still ~400 proteins which are unique to this current study. These may represent potential disease-associated proteins. Among them, several proteins such as serpin B3, renin receptor, and periostin have been reported as pathological markers for renal failure and prostate cancer, respectively. Taken together, our data should provide valuable information for future discovery and validation studies of urine protein biomarkers for various diseases.

60 APPLIED LIFE SCIENCES↗

Prognostic analysis of high-flow nasal cannula therapy and non-invasive ventilation in mild to moderate hypoxemia patients and construction of a machine learning model for 48-h intubation prediction—a retrospective analysis of the MIMIC database

Background This study aims to investigate the clinical outcome between high-flow nasal cannula (HFNC) and non-invasive ventilation (NIV) therapy in mild to moderate hypoxemic patients on the first ICU day and to develop a predictive model of 48-h intubation. Methods The study included adult patients from the MIMIC III and IV databases who first initiated HFNC or NIV therapy due to mild to moderate hypoxemia (100 < PaO2/FiO2 ≤ 300). The 48-h and 30-day intubation rates were compared using cross-sectional and survival analysis. Nine machine learning and six ensemble algorithms were deployed to construct the 48-h intubation predictive models, of which the optimal model was determined by its prediction accuracy. The top 10 risk and protective factors were identified using the Shapley interpretation algorithm. Result A total of 123,042 patients were screened, of which, 673 were from the MIMIC IV database for ventilation therapy comparison (HFNC n = 363, NIV n = 310) and 48-h intubation predictive model construction (training dataset n = 471, internal validation set n = 202) and 408 were from the MIMIC III database for external validation. The NIV group had a lower intubation rate (23.1% vs. 16.1%, p = 0.001), ICU 28-day mortality (18.5% vs. 11.6%, p = 0.014), and in-hospital mortality (19.6% vs. 11.9%, p = 0.007) compared to the HFNC group. Survival analysis showed that the total and 48-h intubation rates were not significantly different. The ensemble AdaBoost decision tree model (internal and external validation set AUROC 0.878, 0.726) had the best predictive accuracy performance. The model Shapley algorithm showed Sequential Organ Failure Assessment (SOFA), acute physiology scores (APSIII), the minimum and maximum lactate value as risk factors for early failure and age, the maximum PaCO 2 and PH value, Glasgow Coma Scale (GCS), the minimum PaO 2 /FiO 2 ratio, and PaO 2 value as protective factors. Conclusion NIV was associated with lower intubation rate and ICU 28-day and in-hospital mortality. Further survival analysis reinforced that the effect of NIV on the intubation rate might partly be attributed to the other impact factors. The ensemble AdaBoost decision tree model may assist clinicians in making clinical decisions, and early organ function support to improve patients’ SOFA, APSIII, GCS, PaCO 2 , PaO 2 , PH, PaO 2 /FiO 2 ratio, and lactate values can reduce the early failure rate and improve patient prognosis.

Fu, Wei↗

Analysis of Spectral Lines in Large Databases of Synthetic Spectra for Massive Stars

In this paper, we describe a program that identifies in the optical spectrum the main parameters of a spectral line, namely the initial and final wavelengths, and the line depth. Moreover, using numerical calculations, it identifies and removes adjacent lines. Next, the program calculates the equivalent width and the FWHM. The software was tested in a sample of 300 lines in two databases of synthetic spectra generated by the CMFGEN and PoWR codes, and 300 lines in observed spectra from the IACOB database, showing a Gaussian distribution of relative errors, from which it is inferred that 80% of the measured lines have errors less than 17% and only 5% of the lines have errors greater than 26%. The program was also run on the entire database of 45,000 CMFGEN and 202 POWR synthetic spectra, generating a library of H i, He i, and He ii lines necessary to feed the FITspec code for the derivation of stellar parameters: effective temperature, surface gravity, and luminosity.

47 OTHER INSTRUMENTATION↗

Linking Extragalactic Transients and Their Host Galaxy Properties: Transient Sample, Multiwavelength Host Identification, and Database Construction

Understanding the preferences of transient types for host galaxies with certain characteristics is key to studies of transient physics and galaxy evolution, as well as to transient identification and classification in the LSST era. Here we describe a value-added database of extragalactic transients—supernovae, tidal disruption events, gamma-ray bursts, and other rare events—and their host galaxy properties. Based on reported coordinates, redshifts, and host galaxies (if known) of events, we cross-identify their host galaxies or most likely host candidates in various value-added or survey catalogs, and compile the existing photometric, spectroscopic, and derived physical properties of the host galaxies in these catalogs. This new database covers photometric measurements from the far-ultraviolet to mid-infrared. Spectroscopic measurements and derived physical properties are also available for a smaller subset of hosts. For our 36,333 unique events, we have cross-identified 13,753 host galaxies using host names, plus 4480 using host coordinates. Besides those with known hosts, there are 18,100 transients with newly identified host candidates. This large database will allow explorations of the connections of transients to their hosts, including a path toward transient alert filtering and probabilistic classification based on host properties.

79 ASTRONOMY AND ASTROPHYSICS↗

The CAI Database: 26 Al– 26 Mg Isotope Systematics

We present a publicly available calcium–aluminum-rich inclusion (CAI) database that focuses on the initial 26 Al/ 27 Al 0 ratio in CAIs, designed in a way that researchers in cosmochemistry and astrophysics may find useful. To date, the database contains 497 CAIs from 75 peer-reviewed papers. The CAIs are from all chondrite groups and cover different CAI types, textures, and sizes. The database includes the paper; the host meteorite; the CAI name and type; the 26 Al/ 27 Al 0 , δ 26 Mg$^*_0$, and δ 25 Mg values and their uncertainties; the number of regression points; the maximum 27 Al/ 24 Mg; the mean-squared weighted deviation; the CAI size; and CAI descriptions. We grouped the CAIs in different ways to discuss 26 Al/ 27 Al 0 ratio distributions with implications for the CAI formation timeline. Overall, we agree with previous authors that CAIs have a bimodal 26 Al distribution: CAIs with robust isochrons (n = 151) have a median 26 Al/ 27 Al 0 = 4.8 × 10 −5 (with a 1σ standard error of 0.1), while those with isotopic anomalies (n = 87) have a median 26 Al/ 27 Al 0 = 0.3 × 10 −5 (with a 1σ standard error of 0.2). However, the large standard deviation of both groups (1.3 and 2.3, respectively) indicates that the 26 Al/ 27 Al 0 values scatter significantly within each population. CAI types and groups can have distinct 26 Al/ 27 Al 0 and δ 26 Mg$^*_0$, but the unmelted inclusions (n = 33) have the highest median 26 Al/ 27 Al 0 = 5.1 × 10 −5 and a low median δ 26 Mg$^*_0$ = −0.05‰. We find slightly different 26 Al/ 27 Al 0 distributions between CAI chondrite types, but no differences between petrographic types or sizes. These observations can help us to understand CAI formation in the context of astrophysical models.

Astronomy and AstroPhysics↗

SoDaH: the SOils DAta Harmonization database, an open-source synthesis of soil data from research networks, version 1.0

Data collected from research networks present opportunities to test theories and develop models about factors responsible for the long-term persistence and vulnerability of soil organic matter (SOM). Synthesizing datasets collected by different research networks presents opportunities to expand the ecological gradients and scientific breadth of information available for inquiry. Synthesizing these data is challenging, especially considering the legacy of soil data that have already been collected and an expansion of new network science initiatives. To facilitate this effort, here we present the SOils DAta Harmonization database (SoDaH; https://lter.github.io/som-website, last access: 22 December 2020), a flexible database designed to harmonize diverse SOM datasets from multiple research networks. SoDaH is built on several network science efforts in the United States, but the tools built for SoDaH aim to provide an open-access resource to facilitate synthesis of soil carbon data. Moreover, SoDaH allows for individual locations to contribute results from experimental manipulations, repeated measurements from long-term studies, and local- to regional-scale gradients across ecosystems or landscapes. Finally, we also provide data visualization and analysis tools that can be used to query and analyze the aggregated database. The SoDaH v1.0 dataset is archived and available at https://doi.org/10.6073/pasta/9733f6b6d2ffd12bf126dc36a763e0b4 (Wieder et al., 2020).

54 ENVIRONMENTAL SCIENCES↗

The ABCflux database: Arctic–boreal CO 2 flux observations and ancillary information aggregated to monthly time steps across terrestrial ecosystems

Past efforts to synthesize and quantify the magnitude and change in carbon dioxide (CO 2 ) fluxes in terrestrial ecosystems across the rapidly warming Arctic–boreal zone (ABZ) have provided valuable information but were limited in their geographical and temporal coverage. Furthermore, these efforts have been based on data aggregated over varying time periods, often with only minimal site ancillary data, thus limiting their potential to be used in large-scale carbon budget assessments. To bridge these gaps, we developed a standardized monthly database of Arctic–boreal CO 2 fluxes (ABCflux) that aggregates in situ measurements of terrestrial net ecosystem CO 2 exchange and its derived partitioned component fluxes: gross primary productivity and ecosystem respiration. The data span from 1989 to 2020 with over 70 supporting variables that describe key site conditions (e.g., vegetation and disturbance type), micrometeorological and environmental measurements (e.g., air and soil temperatures), and flux measurement techniques. Here, we describe these variables, the spatial and temporal distribution of observations, the main strengths and limitations of the database, and the potential research opportunities it enables. In total, ABCflux includes 244 sites and 6309 monthly observations; 136 sites and 2217 monthly observations represent tundra, and 108 sites and 4092 observations represent the boreal biome. The database includes fluxes estimated with chamber (19 % of the monthly observations), snow diffusion (3 %) and eddy covariance (78 %) techniques. The largest number of observations were collected during the climatological summer (June–August; 32 %), and fewer observations were available for autumn (September–October; 25 %), winter (December–February; 18 %), and spring (March–May; 25 %). ABCflux can be used in a wide array of empirical, remote sensing and modeling studies to improve understanding of the regional and temporal variability in CO 2 fluxes and to better estimate the terrestrial ABZ CO 2 budget.

59 BASIC BIOLOGICAL SCIENCES↗

JHTDB-wind: a web-accessible large-eddy simulation database of a wind farm with virtual sensor querying

This paper introduces JHTDB-wind (https://turbulence.idies.jhu.edu/datasets/windfarms, last access: 11 November 2025), a publicly accessible database containing large-eddy simulation (LES) data from wind farms. Building on the framework of the Johns Hopkins Turbulence Database (JHTDB), which hosts direct numerical simulation (DNS) and some LES datasets of canonical turbulent flows, JHTDB-wind stores the 4D space–time history of the flow and provides users the ability to access and query the data via a web-based virtual sensor interface. The initial dataset comprises LES results from a large wind farm with 10×6 turbines, modeled using a filtered actuator line method, under conventionally neutral atmospheric conditions. These data comprise 1 h (hour) of flow field data (velocity, pressure, potential temperature deviation, subgrid-scale (SGS) eddy viscosity, and turbine forces, approximately 15 TB (terabytes) and wind turbine data – including both turbine-level operational quantities and blade-level aerodynamic quantities (approximately 1.3 TB) – stored in Zarr and Parquet formats, respectively. Data retrieval is facilitated by the giverny Python package, allowing remote users to query the database in Python or MATLAB (C and Fortran support are available for flow field data). This paper details the simulation setup and demonstrates data access through examples that analyze wind farm flow structures and turbine performance. The framework is extensible to future datasets, including the JHTDB-wind diurnal cycle simulation analyzed in Xiao et al. (2025).

17 WIND ENERGY↗

Cost considerations in database selection - A comparison of DIALOG and ESA/IRS

It is pointed out that there are many factors which affect the decision-making process in determining which databases should be selected for conducting the online search on a given topic. In many cases, however, the major consideration will be related to cost. The present investigation is concerned with a comparison of the costs involved in making use of DIALOG and the European Space Agency's Information Retrieval Service (ESA/IRS). The two services are very comparable in many respects. Attention is given to pricing structure, telecommunications, the number of databases, prints, time requirements, a table listing online costs for DIALOG and ESA/IRS, and differences in mounting databases. It is found that ESA/IRS is competitively priced when compared to DIALOG, and, despite occasionally higher telecommunications costs, may be even more economical to use in some cases.

Jack, R. F.↗

The use of database management systems and artificial intelligence in automating the planning of optical navigation pictures

The use of database management systems (DBMS) and AI to minimize human involvement in the planning of optical navigation pictures for interplanetary space probes is discussed, with application to the Galileo mission. Parameters characterizing the desirability of candidate pictures, and the program generating them, are described. How these parameters automatically build picture records in a database, and the definition of the database structure, are then discussed. The various rules, priorities, and constraints used in selecting pictures are also described. An example is provided of an expert system, written in Prolog, for automatically performing the selection process.

Davis, Robert P.↗

Automated database design technology and tools

The Automated Database Design Technology and Tools research project results are summarized in this final report. Comments on the state of the art in various aspects of database design are provided, and recommendations made for further research for SNAP and NAVMASSO future database applications.

Shen, Stewart N. T.↗

The HITRAN database - 1986 edition

A description and summary of the latest edition of the AFGL high-resolution transmission molecular absorption database (HITRAN) parameters are presented. This new database combines the information for the seven principal atmospheric absorbers and twenty-one additional molecular species previously contained on the AFGL atmospheric absorption line parameter compilation and on the trace gas compilation. In addition to updating the parameters on earlier editions of the compilation, new parameters have been added to this edition such as the self-broadened half-width, the temperature dependence of the air-broadened half-width, and the transition probability. The database contains 348,043 entries between 0 and 17,900/cm. A FORTRAN program is now furnished to allow rapid access to the molecular transitions and for the creation of customized output. A separate file of molecular cross sections of 11 heavy molecular species, applicable for qualitative simulation of transmission and emission in the atmosphere, has also been provided.

Rothman, L. S.↗

An automated system for terrain database construction

An automated Terrain Database Preparation System (TDPS) for the construction and editing of terrain databases used in computerized wargaming simulation exercises has been developed. The TDPS system operates under the TAE executive, and it integrates VICAR/IBIS image processing and Geographic Information System software with CAD/CAM data capture and editing capabilities. The terrain database includes such features as roads, rivers, vegetation, and terrain roughness.

Johnson, L. F.↗

The Einstein observatory stellar X-ray database - An overview

The paper presents the motivations for and the methodology followed in building the 'Einstein Observatory Stellar X-ray Database' based on the uniform analysis of all Einstein Observatory Imaging Proportional Counter fields obtained during the life of the HEAO-2 mission. The database has been implemented using the INGRES database system, so that statistical analyses of the properties of the full detection catalog are relatively easily and flexibly accomplished. Some illustrative examples will furnish a general view both of the kind and the amount of the archived information, and of the statistical approach used in analyzing the global properties of the data.

Sciortino, S.↗

Using decision-tree classifier systems to extract knowledge from databases

One difficulty in applying artificial intelligence techniques to the solution of real world problems is that the development and maintenance of many AI systems, such as those used in diagnostics, require large amounts of human resources. At the same time, databases frequently exist which contain information about the process(es) of interest. Recently, efforts to reduce development and maintenance costs of AI systems have focused on using machine learning techniques to extract knowledge from existing databases. Research is described in the area of knowledge extraction using a class of machine learning techniques called decision-tree classifier systems. Results of this research suggest ways of performing knowledge extraction which may be applied in numerous situations. In addition, a measurement called the concept strength metric (CSM) is described which can be used to determine how well the resulting decision tree can differentiate between the concepts it has learned. The CSM can be used to determine whether or not additional knowledge needs to be extracted from the database. An experiment involving real world data is presented to illustrate the concepts described.

St.clair, D. C.↗

Geometric database maintenance using CCTV cameras and overlay graphics

An interactive graphics system using closed circuit television (CCTV) cameras for remote verification and maintenance of a geometric world model database has been demonstrated in GE's telerobotics testbed. The database provides geometric models and locations of objects viewed by CCTV cameras and manipulated by telerobots. To update the database, an operator uses the interactive graphics system to superimpose a wireframe line drawing of an object with known dimensions on a live video scene containing that object. The methodology used is multipoint positioning to easily superimpose a wireframe graphic on the CCTV image of an object in the work scene. An enhanced version of GE's interactive graphics system will provide the object designation function for the operator control station of the Jet Propulsion Laboratory's telerobot demonstration system.

Oxenberg, Sheldon C.↗