Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “hierarchical sampling”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Bayesian Statistical Models for Community Annoyance Survey Data

This paper demonstrates the use of two Bayesian statistical models to analyze single-event sonic boom exposure and human annoyance data from community response surveys. Each model is fit to data from a NASA pilot study.Unlike many community noise surveys, this study used a panel sample to collect multiple observations per participant instead of a single observation. Thus, a multilevel (also known as hierarchical or mixed-effects) model is used to account for the within-subject correlation in the panel sample data. This paper describes a multilevel logistic regression model and a multilevel ordinal regression model. The paper also proposes a method for calculating a summary dose-response curve from the multilevel models that represents the population. The two models’ summary dose-response curves are visually similar. However, their estimates differ when calculating the noise dose at a fixed percent highly annoyed.

Musical instruments↗

Harvest Initiated Volatile Organic Compound Emissions from In-Field Tall Wheatgrass

While crop and grassland usage continues to increase, the full diversity of plant-specific volatile organic compounds (VOCs) emitted from these ecosystems, including their implications for atmospheric chemistry and carbon cycling, remains poorly understood. It is particularly important to investigate VOCs in the context of potential biofuels: aside from the implications of largescale land use, harvest may shift both the flux and speciation of emitted VOCs. To this point, we evaluate the diversity of VOCs emitted both pre and postharvest from “Alkar” tall wheatgrass (Thinopyrum ponticum), a candidate biofuel that exhibits greater tolerance to frost and saline land compared to other grass varieties. Mature plants grown under field conditions (n = 6) were sampled for VOCs both pre- and postharvest (October 2022). Via hierarchical clustering of emitted VOCs from each plant, we observe distinct “volatilomes” (diversity of VOCs) specific to the pre- and postharvest conditions despite plant-to-plant variability. In total, 50 VOCs were found to be unique to the postharvest tall wheatgrass volatilome, and these unique VOCs constituted a significant portion (26%) of total postharvest signal. While green leaf volatiles (GLVs) dominate the speciation of postharvest emissions (e.g., 54% of unique postharvest VOC signal was due to 1-penten-3-ol), we demonstrate novel postharvest VOCs from tall wheatgrass that are under characterized in the context of carbon cycling and atmospheric chemistry (e.g., 3-octanone). Continuing evaluations will quantitatively investigate tall wheatgrass VOC fluxes, better informing the feasibility and environmental impact of tall wheatgrass as a biofuel.

09 BIOMASS FUELS↗

Hierarchical screening for Li-based solid electrolytes using fast, interpretable machine-learned potentials

Li-based solid-state electrolyte materials enable safer, all-solid-state batteries but the computational search for candidates with favorable stability and Li-ion conductivity is challenging due to the size of the search space and the cost of evaluating transport properties with ab initio methods. The prohibitive cost of high-throughput screening with DFT has lead to the development of surrogate models using geometric analysis, empirical potentials, and descriptors for ionic transport. Here, I will discuss a hierarchical screening approach for identifying promising materials using a combination of density functional theory, bond-valence methods, and machine learning potentials generated with the Ultra-Fast Force Fields (UF3) framework. We show how the inexpensive bond-valence method can be used to guide the generation of training samples for machine learning, in addition to filtering candidates. Finally, we apply the hierarchical workflow to screen for ionic conductivity across a database of Li-containing compounds.

Materials discovery↗

The effects of cloud inhomogeneities upon radiative fluxes, and the supply of a cloud truth validation dataset

The ASTER polar cloud mask algorithm is currently under development. Several classification techniques have been developed and implemented. The merits and accuracy of each are being examined. The classification techniques under investigation include fuzzy logic, hierarchical neural network, and a pairwise histogram comparison scheme based on sample histograms called the Paired Histogram Method. Scene adaptive methods also are being investigated as a means to improve classifier performance. The feature, arctan of Band 4 and Band 5, and the Band 2 vs. Band 4 feature space are key to separating frozen water (e.g., ice/snow, slush/wet ice, etc.) from cloud over frozen water, and land from cloud over land, respectively. A total of 82 Landsat TM circumpolar scenes are being used as a basis for algorithm development and testing. Numerous spectral features are being tested and include the 7 basic Landsat TM bands, in addition to ratios, differences, arctans, and normalized differences of each combination of bands. A technique for deriving cloud base and top height is developed. It uses 2-D cross correlation between a cloud edge and its corresponding shadow to determine the displacement of the cloud from its shadow. The height is then determined from this displacement, the solar zenith angle, and the sensor viewing angle.

Welch, Ronald M.↗

Exploring the energy landscape of RBMs: reciprocal space insights into bosons, hierarchical learning and symmetry breaking

Deep generative models have become ubiquitous due to their ability to learn and sample from complex distributions. Despite the proliferation of various frameworks, the relationships among these models remain largely unexplored, a gap that hinders the development of a unified theory of AI learning. In this work, we address two central challenges: clarifying the connections between different deep generative models and deepening our understanding of their learning mechanisms. We focus on Restricted Boltzmann Machines (RBMs), a class of generative models known for their universal approximation capabilities for discrete distributions. By introducing a reciprocal space formulation for RBMs, we reveal a connection between these models, diffusion processes, and systems of coupled bosons. Our analysis shows that at initialization, the RBM operates at a saddle point, where the local curvature is determined by the singular values of the weight matrix, whose distribution follows the Marc̆enko-Pastur law and exhibits rotational symmetry. During training, this rotational symmetry is broken due to hierarchical learning, where different degrees of freedom progressively capture features at multiple levels of abstraction. This leads to a symmetry breaking in the energy landscape, reminiscent of Landau’s theory. This symmetry breaking in the energy landscape is characterized by the singular values and the weight matrix eigenvector matrix. We derive the corresponding free energy in a mean-field approximation. We show that in the limit of infinite size RBM, the reciprocal variables are Gaussian distributed. Our findings indicate that in this regime, there will be some modes for which the diffusion process will not converge to the Boltzmann distribution. To illustrate our results, we trained replicas of RBMs with different hidden layer sizes using the MNIST dataset. Our findings not only bridge the gap between disparate generative frameworks but also shed light on the fundamental processes underpinning learning in deep generative models.

97 MATHEMATICS AND COMPUTING↗

The eROSITA Final Equatorial-Depth Survey (eFEDS): The AGN catalog and its X-ray spectral properties

The eROSITA Final Equatorial Depth Survey (eFEDS), observed with eROSITA ahead of its planned 4-yr all-sky survey, is the largest contiguous-field X-ray survey at present. It yielded a large sample of X-ray sources with very rich multiband photometric and spectroscopic coverage. We present here the eFEDS active galactic nuclei (AGN) catalog and the eROSITA X-ray spectral properties of the eFEDS sources. Using a Bayesian method, we performed a systematic X-ray spectral analysis for all the eFEDS sources. We adopted multiple spectral models, including single-component power-law or hot-plasma models and double-component models of a power law plus soft excess. We investigated the capacity of eROSITA X-ray spectra for constraining AGN spectral shapes through a detailed analysis of the posterior parameter probability distribution functions. Hierarchical Bayesian modeling was used to recover the spectral parameter distribution of the sample. The source fluxes and luminosities were measured from the posterior of the spectral fitting. The eFEDS AGN catalog (22 079 sources) comprises ~80% of the eFEDS point sources. Despite a large number of faint sources, our spectral fitting provides reasonable measurements of spectral shapes and intrinsic luminosities for a majority of the sources. Because of sample selection bias, this AGN catalog is dominated by X-ray unobscured sources, with an obscured (logN H > 21.5) fraction of 8%; the power-law emission of the hot corona is also relatively soft, with a typical slope of 2.0. For type-I AGN, the X-ray emission is well correlated with the UV emission with the usual anticorrelation between the X-ray to UV spectral slope α OX and the UV luminosity. The X-ray spectral properties measured with various models are presented for all the eFEDS sources.

79 ASTRONOMY AND ASTROPHYSICS↗

Mars Science Laboratory CHIMRA/IC/DRT Flight Software for Sample Acquisition and Processing

The design methodologies of using sequence diagrams, multi-process functional flow diagrams, and hierarchical state machines were successfully applied in designing three MSL (Mars Science Laboratory) flight software modules responsible for handling actuator motions of the CHIMRA (Collection and Handling for In Situ Martian Rock Analysis), IC (Inlet Covers), and DRT (Dust Removal Tool) mechanisms. The methodologies were essential to specify complex interactions with other modules, support concurrent foreground and background motions, and handle various fault protections. Studying task scenarios with multi-process functional flow diagrams yielded great insight to overall design perspectives. Since the three modules require three different levels of background motion support, the methodologies presented in this paper provide an excellent comparison. All three modules are fully operational in flight.

sample processing↗

TDCOSMO XXIII. Measurement of the Hubble constant from the doubly lensed quasar HE 1104−1805

Time-delay cosmography leverages strongly lensed quasars to measure the Universe’s current expansion rate, H 0 , independently from other methods. The latest TDCOSMO milestone measurement primarily used quadruply lensed quasars for their mass profile constraints. However, doubly lensed quasars, being more abundant and offering precise time delays, could expand the sample by a factor of 5, significantly advancing towards a 1% precision measurement of H 0 . We present the first TDCOSMO analysis of a doubly imaged source, HE 1104−1805, including the measurement of the four necessary ingredients. First, by combining 17 years of data from the SMARTS, Euler, and WFI telescopes, we measured a time delay of 176.3 +11.4 −10.3 days. Second, using MUSE data, we extracted stellar velocity dispersion measurements in three radial bins with 5% to 13% precision. Third, employing F160W HST imaging for lens modelling and marginalising over various modelling choices, we measured the Fermat potential difference between the images. Fourth, using wide-field imaging, we measured the convergence added by objects not included in the lens modelling. By combining these four ingredients, we measured the time delay distance and the angular diameter distance to the deflector, favouring a power-law mass model over a baryonic and dark matter composite model. The measurement was performed blindly to prevent experimenter bias and resulted in a Hubble constant of H 0 = +5.8 −5.0 × λ int km s −1 Mpc −1 , where λ int is the internal mass sheet degeneracy parameter. This is in agreement with the TDCOSMO-2025 milestone and its precision for λ int = 1 is comparable to that obtained with the best-observed quadruply lensed quasars (4–6%). This work is a stepping stone towards a precise measurement of H 0 using a large sample of doubly lensed quasars, supplementing the current sample. The next TDCOSMO milestone paper will include this system in its hierarchical analysis, constraining λ int and H 0 jointly with multiple lenses.

cosmological parameters↗

The extended medium sensitivity survey distant cluster sample - X-ray data and interpretation of the luminosity evolution

The X-ray properties of a cluster of galaxies subsample of the Einstein Extended Medium Sensitivity Survey is described. A summary of this sample and its implication has been presented previously; this paper gives the full details. The cluster subsample is 98.4 percent identified and contains 93 X-ray-selected clusters to a redshift of 0.58. The cluster X-ray luminosity function at three cosmic epochs is derived. While the present luminosity function agrees with previous determinations at the lowest redshifts, it is found that the volume density of high-luminosity clusters is greater now than it was in the past. The normalization, shape, and time dependence of the luminosity function can be described by a simple hierarchical formation model with parameters which also describe the temperature function of an independent sample of low-redshift clusters. In this model the comoving hot gas density remains constant with time at least to redshifts of order 0.35.

Henry, J. P.↗

TDCOSMO VIII. A key test of systematics in the hierarchical method of time-delay cosmography

The largest source of systematic errors in the time-delay cosmography method likely arises from the lens model mass distribution, where an inaccurate choice of model could in principle bias the value of $H$ 0 . A Bayesian hierarchical framework has been proposed which combines lens systems with kinematic data, constraining the mass profile shape at a population level. The framework has been previously validated using a small sample of lensing galaxies drawn from hydro-simulations. The goal of this work is to expand the validation to a more general set of lenses consistent with observed systems, as well as confirm the capacity of the method to combine two lens populations: one which has time delay information and one which lacks time delays and has systematically different image radii. For this purpose, we generated samples of analytic lens mass distributions made of baryons+dark matter and fit the subsequent mock images with standard power-law models. Corresponding kinematics data were also emulated. The hierarchical framework applied to an ensemble of time-delay lenses allowed us to correct the $H$ 0 bias associated with model choice to find $H$ 0 within 1.5σ of the fiducial value. We then combined this set with a sample of corresponding lens systems which have no time delays and have a source at lower $z$, resulting in a systematically smaller image radius relative to their effective radius. The hierarchical framework has successfully accounted for this effect, recovering a value of $H$ 0 which is both more precise (σ ~ 2%) and more accurate (0.7% median offset) than the time-delay set alone. This result confirms that non-time-delay lenses can nonetheless contribute valuable constraining power to the determination of $H$ 0 via their kinematic constraints, assuming they come from the same global population as the time-delay set.

79 ASTRONOMY AND ASTROPHYSICS↗

An Introduction to the Federated Architecture for Secure and Transactive Distributed Energy Management Solutions (FAST-DERMS)

Deployment and capability of distributed energy resources (DER) in power systems is growing rapidly. These resources present an opportunity for low-cost provision of energy and grid services. The Federal Energy Regulatory Commission recently provided rulings to enable market participation of these distribution-connected resources, but the prevailing strategies for their management may not scale well to meet future needs. This paper introduces the Federated Architecture for Secure and Transactive Distributed Energy Management Solutions (FAST-DERMS) which was designed to address this need. In it we describe the architectural features of the approach, and a reference controls implementation employing a hierarchical coordination that includes stochastic optimization, model predictive control, and a simple real-time management scheme. Sample results from simulation show firm transmission-level service provision measured at the distribution substation.

grid architecture↗

An Introduction to the Federated Architecture for Secure and Transactive Distributed Energy Management Solutions (FAST-DERMS): Preprint

Deployment and capability of distributed energy resources (DER) in power systems is growing rapidly. These resources present an opportunity for low-cost provision of energy and grid services. The Federal Energy Regulatory Commission recently provided rulings to enable market participation of these distribution-connected resources, but the prevailing strategies for their management may not scale well to meet future needs. This paper introduces the Federated Architecture for Secure and Transactive Distributed Energy Management Solutions (FASTDERMS) which was designed to address this need. In it we describe the architectural features of the approach, and a reference controls implementation employing a hierarchical coordination that includes stochastic optimization, model predictive control, and a simple real-time management scheme. Sample results from simulation show firm transmission-level service provision measured at the distribution substation.

DERMS↗

Accelerating Random Forest Classification on GPU and FPGA

Random Forests (RFs) are a commonly used machine learning method for classification and regression tasks spanning a variety of application domains, including bioinformatics, business analytics, and software optimization. While prior work has focused primarily on improving performance of the training of RFs, many applications, such as malware identification, cancer prediction, and banking fraud detection, require fast RF classification. In this work, we accelerate RF classification on GPU and FPGA. In order to provide efficient support for large datasets, we propose a hierarchical memory layout suitable to the GPU/FPGA memory hierarchy. We design three RF classification code variants based on that layout, and we investigate GPU- and FPGA-specific considerations for these kernels. Our experimental evaluation, performed on an Nvidia Xp GPU and on a Xilinx Alveo U250 FPGA accelerator card using publicly available datasets on the scale of millions of samples and tens of features, covers various aspects. First, we evaluate the performance benefits of our hierarchical data structure over the standard compressed sparse row (CSR) format. Second, we compare our GPU implementation with cuML, a machine learning library targeting Nvidia GPUs. Third, we explore the performance/accuracy tradeoff resulting from the use of different tree depths in the RF. Finally, we perform a comparative performance analysis of our GPU and FPGA implementations. Our evaluation shows that for high accuracy targets, our GPU implementation yields 5-9x speedup over CSR, and up to a 2x speedup over cuML.

FPGA, Xilinx FPGA, GPU, Random Forest classificati↗

Unified many-worlds browsing of arbitrary physics-based animations

Manually tuning physics-based animation parameters to explore a simulation outcome space or achieve desired motion outcomes can be notoriously tedious. This problem has motivated many sophisticated and specialized optimization-based methods for fine-grained (keyframe) control, each of which are typically limited to specific animation phenomena, usually complicated, and, unfortunately, not widely used. In this paper, we propose Unified Many-Worlds Browsing (UMWB), a practical method for sample-level control and exploration of physics-based animations. Our approach supports browsing of large simulation ensembles of arbitrary animation phenomena by using a unified volumetric WORLDPACK representation based on spatiotemporally compressed voxel data associated with geometric occupancy and other low-fidelity animation state. Beyond memory reduction, the WORLDPACK representation also enables unified query support for interactive browsing: it provides fast evaluation of approximate spatiotemporal queries, such as occupancy tests that find ensemble samples ("worlds") where material is either IN or NOT IN a user-specified spacetime region. WORLDPACKS also support real-time hardware-accelerated voxel rendering by exploiting the spatially hierarchical and temporal RLE raster data structure. Our UMWB implementation supports interactive browsing (and offline refinement) of ensembles containing thousands of simulation samples, and fast spatiotemporal queries and ranking. We show UMWB results using a wide variety of physics-based animation phenomena---not just JELL-O ® .

Computer Science↗

Substructure in the stellar halo near the Sun: I. Data-driven clustering in integrals-of-motion space

Context. Merger debris is expected to populate the stellar haloes of galaxies. In the case of the Milky Way, this debris should be apparent as clumps in a space defined by the orbital integrals of motion of the stars. Aims. Our aim is to develop a data-driven and statistics-based method for finding these clumps in integrals-of-motion space for nearby halo stars and to evaluate their significance robustly. Methods. We used data from Gaia EDR3, extended with radial velocities from ground-based spectroscopic surveys, to construct a sample of halo stars within 2.5 kpc from the Sun. We applied a hierarchical clustering method that makes exhaustive use of the single linkage algorithm in three-dimensional space defined by the commonly used integrals of motion energy E, together with two components of the angular momentum, L z and L ⊥ . To evaluate the statistical significance of the clusters, we compared the density within an ellipsoidal region centred on the cluster to that of random sets with similar global dynamical properties. By selecting the signal at the location of their maximum statistical significance in the hierarchical tree, we extracted a set of significant unique clusters. By describing these clusters with ellipsoids, we estimated the proximity of a star to the cluster centre using the Mahalanobis distance. Additionally, we applied the HDBSCAN clustering algorithm in velocity space to each cluster to extract subgroups representing debris with different orbital phases. Results. Our procedure identifies 67 highly significant clusters (> 3σ), containing 12% of the sources in our halo set, and 232 subgroups or individual streams in velocity space. In total, 13.8% of the stars in our data set can be confidently associated with a significant cluster based on their Mahalanobis distance. Inspection of the hierarchical tree describing our data set reveals a complex web of relations between the significant clusters, suggesting that they can be tentatively grouped into at least six main large structures, many of which can be associated with previously identified halo substructures, and a number of independent substructures. This preliminary conclusion is further explored in a companion paper, in which we also characterise the substructures in terms of their stellar populations. Conclusions. Our method allows us to systematically detect kinematic substructures in the Galactic stellar halo with a data-driven and interpretable algorithm. The list of the clusters and the associated star catalogue are provided in two tables available at the CDS.

79 ASTRONOMY AND ASTROPHYSICS↗

DAmodel: hierarchical Bayesian modelling of DA white dwarfs for spectrophotometric calibration

We use hierarchical Bayesian modelling to calibrate a network of 32 all-sky faint DA white dwarf (DA WD) spectrophotometric standards (⁠16.5 < V , 19.5⁠) alongside three CALSPEC standards, from 912 Å to 32 μm. The framework is the first of its kind to jointly infer photometric zero points and WD parameters (surface gravity log g⁠, effective temperature T eff ⁠, extinction A V ⁠, dust relation parameter R V ) by simultaneously modelling both photometric and spectroscopic data. We model panchromatic Hubble Space Telescope Wide Field Camera 3 (HST/WFC3) UVIS and IR photometry, HST/STIS UV spectroscopy, and ground-based optical spectroscopy to sub-per cent precision. Photometric residuals for the sample are the lowest yet yielding < 0.004 mag RMS on average from the UV to the NIR, achieved by jointly inferring time-dependent changes in system sensitivity and WFC3/IR count-rate nonlinearity. Our GPU-accelerated implementation enables efficient sampling via Hamiltonian Monte Carlo, critical for exploring the high-dimensional posterior space. The hierarchical nature of the model enables population analysis of intrinsic WD and dust parameters. Inferred spectral energy distributions from this model will be essential for calibrating the James Webb Space Telescope as well as next-generation surveys, including Vera Rubin Observatory’s Legacy Survey of Space and Time and the Nancy Grace Roman Space Telescope.

methods: statistical↗

Analysis of interstellar cloud structure based on IRAS images

The goal of this project was to develop new tools for the analysis of the structure of densely sampled maps of interstellar star-forming regions. A particular emphasis was on the recognition and characterization of nested hierarchical structure and fractal irregularity, and their relation to the level of star formation activity. The panoramic IRAS images provided data with the required range in spatial scale, greater than a factor of 100, and in column density, greater than a factor of 50. In order to construct densely sampled column density maps of star-forming clouds, column density images of four nearby cloud complexes were constructed from IRAS data. The regions have various degrees of star formation activity, and most of them have probably not been affected much by the disruptive effects of young massive stars. The largest region, the Scorpius-Ophiuchus cloud complex, covers about 1000 square degrees (it was subdivided into a few smaller regions for analysis). Much of the work during the early part of the project focused on an 80 square degree region in the core of the Taurus complex, a well-studied region of low-mass star formation.

Scalo, John M.↗

Analysis of interstellar fragmentation structure based on IRAS images

The goal of this project was to develop new tools for the analysis of the structure of densely sampled maps of interstellar star-forming regions. A particular emphasis was on the recognition and characterization of nested hierarchical structure and fractal irregularity, and their relation to the level of star formation activity. The panoramic IRAS images provided data with the required range in spatial scale, greater than a factor of 100, and in column density, greater than a factor of 50. In order to construct a densely sampled column density map of a cloud complex which is both self-gravitating and not (yet?) stirred up much by star formation, a column density image of the Taurus region has been constructed from IRAS data. The primary drawback to using the IRAS data for this purpose is that it contains no velocity information, and the possible importance of projection effects must be kept in mind.

Scalo, John M.↗