Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “ensemble clustering”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Predictions of Gamma-ray Emission from Globular Cluster Millisecond Pulsars Above 100 MeV

The recent Fermi detection of the globular cluster (GC) 47 Tucanae highlighted the importance of modeling collective gamma-ray emission of millisecond pulsars (MSPs) in GCs. Steady flux from such populations is also expected in the very high energy (VHE) domain covered by ground-based Cherenkov telescopes. We present pulsed curvature radiation (CR) as well as unpulsed inverse Compton (IC) calculations for an ensemble of MSPs in the GCs 47 Tucanae and Terzan 5. We demonstrate that the CR from these GCs should be easily detectable for Fermi, while constraints on the total number of MSps and the nebular B-field may be derived using the IC flux components.

Venter, C.↗

A preliminary study of the tropical water cycle and its sensitivity to surface warming

The Goddard Cumulus Ensemble Model (GCEM) has been used to demonstrate that cumulus-scale dynamics and microphysics play a major role in determining the vertical distribution of water vapor and clouds in the tropical atmosphere. The GCEM is described and is the basic structure of cumulus convection. The long-term equilibrium response to tropical convection to surface warming is examined. A picture of the water cycle within tropical cumulus clusters is developed.

Lau, K. M.↗

T-Matrix Modeling of Linear Depolarization by Morphologically Complex Soot and Soot-Containing Aerosols

We use state-of-the-art public-domain Fortran codes based on the T-matrix method to calculate orientation and ensemble averaged scattering matrix elements for a variety of morphologically complex black carbon (BC) and BC-containing aerosol particles, with a special emphasis on the linear depolarization ratio (LDR). We explain theoretically the quasi-Rayleigh LDR peak at side-scattering angles typical of low-density soot fractals and conclude that the measurement of this feature enables one to evaluate the compactness state of BC clusters and trace the evolution of low-density fluffy fractals into densely packed aggregates. We show that small backscattering LDRs measured with groundbased, airborne, and spaceborne lidars for fresh smoke generally agree with the values predicted theoretically for fluffy BC fractals and densely packed near-spheroidal BC aggregates. To reproduce higher lidar LDRs observed for aged smoke, one needs alternative particle models such as shape mixtures of BC spheroids or cylinders.

atmospheric radiation↗

Evaluation of Clouds, Radiation, and Precipitation in CMIP6 Models Using Global Weather States Derived from ISCCP-H Cloud Property Data

A clustering methodology is applied to cloud optical depth cloud top pressure (TAU-PC) histograms from the new, 1-degree resolution, ISCCP-H dataset, to derive an updated global Weather State (WS) dataset. Then, PC-TAU histograms from current-climate CMIP6 model simulations are assigned to the ISCCP-H WSs along with their concurrent radiation and precipitation properties, to evaluate model cloud, radiation, and precipitation properties in the context of the Weather States. The new ISCCP-H analysis produces WSs that are very similar to those previously found in the lower resolution ISCCP-D dataset. The main difference lies in the splitting of the ISCCP-D thin stratocumulus WS between the ISCCP-H shallow cumulus and stratocumulus WSs, which results in the reduction by one of the total WS number. The evaluation of the CMIP6 models against the ISCCP-H Weather States, shows that, in the ensemble mean, the models are producing an adequate representation of the frequency and geographical distribution of the WSs, with measurable improvements compared to the WSs derived for the CMIP5 ensemble. However, the frequency of shallow cumulus clouds continues to be underestimated, and, in some WSs the good agreement of the ensemble mean with observations comes from averaging models that significantly overpredict and underpredict the ISCCP-H WS frequency. In addition, significant biases exist in the internal cloud properties of the model WSs, such as the model underestimation of cloud fraction in middle-top clouds and secondarily in midlatitude storm and stratocumulus clouds, that result in an underestimation of cloud SW cooling in those regimes.

Clouds↗

Extracting Independent Local Oscillatory Geophysical Signals by Geodetic Tropospheric Delay

Zenith Tropospheric Delay (ZTD) due to water vapor derived from space geodetic techniques and numerical weather prediction simulated-reanalysis data exhibits non-linear and non-stationary properties akin to those in the crucial geophysical signals of interest to the research community. These time series, once decomposed into additive (and stochastic) components, have information about the long term global change (the trend) and other interpretable (quasi-) periodic components such as seasonal cycles and noise. Such stochastic component(s) could be a function that exhibits at most one extremum within a data span or a monotonic function within a certain temporal span. In this contribution, we examine the use of the combined Ensemble Empirical Mode Decomposition (EEMD) and Independent Component Analysis (ICA): the EEMD-ICA algorithm to extract the independent local oscillatory stochastic components in the tropospheric delay derived from the European Centre for Medium-Range Weather Forecasts (ECMWF) over six geodetic sites (HartRAO, Hobart26, Wettzell, Gilcreek, Westford, and Tsukub32). The proposed methodology allows independent geophysical processes to be extracted and assessed. Analysis of the quality index of the Independent Components (ICs) derived for each cluster of local oscillatory components (also called the Intrinsic Mode Functions (IMFs)) for all the geodetic stations considered in the study demonstrate that they are strongly site dependent. Such strong dependency seems to suggest that the localized geophysical signals embedded in the ZTD over the geodetic sites are not correlated. Further, from the viewpoint of non-linear dynamical systems, four geophysical signals the Quasi-Biennial Oscillation (QBO) index derived from the NCEP/NCAR reanalysis, the Southern Oscillation Index (SOI) anomaly from NCEP, the SIDC monthly Sun Spot Number (SSN), and the Length of Day (LoD) are linked to the extracted signal components from ZTD. Results from the synchronization analysis show that ZTD and the geophysical signals exhibit (albeit subtle) site dependent phase synchronization index.

Botai, O. J.↗

NASA’s Prototype Spectral Water Inversion Processor and Emulator (SWIPE): Towards Global Coastal and Inland Water Quality and Algal Biodiversity Monitoring

Degradation of Earth’s inland water resources due to anthropogenic perturbations and climate anomalies at both local and global scales continues to place human health at substantial risk. There is now a growing necessity to develop pragmatic approaches that allow timely and effective extrapolation of local processes, to spatially resolved global products, and to promote operational and sustainable resource policy management. This presentation will provide updates on NASA’s prototype open-source aquatic modeling platform, Spectral Water Inversion Processor and Emulator (SWIPE), which is a comprehensive, multi-faceted modeling platform for both forward and inverse modeling of diverse aquatic ecosystems from the benthos to top-of-atmosphere (TOA). SWIPE provides a cohesive application which leverages recent advancements in particle modeling, Big Data analytics, and machine learning to develop a high-fidelity synthetic training ground for sensitivity studies and algorithm development for multispectral or upcoming hyperspectral missions. Some of the prominent features of SWIPE to be discussed include: 1. Advanced hyperspectral modeling of globally diverse algal and non-algal particles using a novel two-layer coated sphere scattering model and radiative transfer modeling, 2. Massive, highly detailed synthetic spectral libraries of Analysis-Ready-Data (ARD) which include spectral libraries of particle microphysics, water biogeophysical and optical properties, as well as surface and TOA reflectances at 1 nm resolution, 3. An ensemble of pre-built analytic, machine learning, and deep learning inversion algorithms for various water quality and biodiversity related retrieval parameters and uncertainty quantification, 4. Sensor-agnostic water quality inversion at wide ranging spatial and spectral resolutions including a codebase for seamless application in the Google Earth Engine and NASA Earth Exchange (NEX) for planetary scale analysis. SWIPE will be a fully open-source platform based in python with comprehensive documentation, tutorials, and options for distributed computing on high performance computing clusters or on single, local machines. Further, we will discuss how we envision SWIPE contributing towards a global analysis of coastal and inland water quality dynamics.

top-of-atmosphere (TOA)↗

Constraints on dark matter from cosmic background anisotropies

The major stages in the linear evolution of the statistical ensemble of adiabatic fluctuations in radiation, baryons, and dark matter are discussed. If it is assumed that the distribution of light emitters (i.e., of galaxies) follows the distribution of mass (i.e., of dark matter), then universes dominated by massive collisionless relics of the Big Bang must have Omega larger than 0.2 h exp -4/3 to avoid exceeding the current observational limits on small-scale anisotropies in the microwave background. However, values of Omega of about 0.2 are indicated by dynamical studies of galaxy clustering. It is concluded that universes dominated by cold dark matter in which light traces mass are probably not viable models.

Bond, J. R.↗

A CO survey of regions around 34 open clusters. II - Physical properties of cataloged molecular clouds

The physical properties of the 148 molecular clouds found in a CO survey of regions around 34 young open clusters have been examined. Expressions are given for the cloud size spectrum and the mass spectrum. The mass-radius relation implies that clouds of all size larger than a few pc have about the same mean volume density. Power laws with slopes of 0.6 and 3 describe, respectively, the relations of CO linewidth and cloud mass to cloud size. The clouds are distinctly nonspherical and appear to be randomly oriented with respect to the Galactic plane. The observations can be explained by a model for molecular clouds in which clouds are ensembles of dense clumps of gas. Based on such a model, it is shown that molecular clouds are perturbed on a time scale short compared to the time required for them to reestablish virial equilibrium.

Leisawitz, D.↗

High Speed Networking and Large-scale Simulation in Geodynamics

Large-scale numerical simulation has been one of the most important approaches for understanding global geodynamical processes. In this approach, peta-scale floating point operations (pflops) are often required to carry out a single physically-meaningful numerical experiment. For example, to model convective flow in the Earth's core and generation of the geomagnetic field (geodynamo), simulation for one magnetic free-decay time (approximately 15000 years) with a modest resolution of 150 in three spatial dimensions would require approximately 0.2 pflops. If such a numerical model is used to predict geomagnetic secular variation over decades and longer, with e.g. an ensemble Kalman filter assimilation approach, approximately 30 (and perhaps more) independent simulations of similar scales would be needed for one data assimilation analysis. Obviously, such a simulation would require an enormous computing resource that exceeds the capacity of a single facility currently available at our disposal. One solution is to utilize a very fast network (e.g. 10Gb optical networks) and available middleware (e.g. Globus Toolkit) to allocate available but often heterogeneous resources for such large-scale computing efforts. At NASA GSFC, we are experimenting with such an approach by networking several clusters for geomagnetic data assimilation research. We shall present our initial testing results in the meeting.

Kuang, Weijia↗

Control of the NASA Langley 16-Foot Transonic Tunnel with the Self-Organizing Map

A predictive, multiple model control strategy is developed based on an ensemble of local linear models of the nonlinear system dynamics for a transonic wind tunnel. The local linear models are estimated directly from the weights of a self-organizing map (SOM). Multiple self-organizing maps collectively model the global response of the wind tunnel to a finite set of representative prototype controls. These prototype controls partition the control space and incorporate experiential knowledge gained from decades of operation. Each SOM models the combination of the tunnel with one of the representative controls, over the entire range of operation. The SOM based linear models are used to predict the tunnel response to a larger family of control sequences which are clustered on the representative prototypes. The control sequence which corresponds to the prediction that best satisfies the requirements on the system output is applied as the external driving signal.

Mark A. Motter↗

Regional and Model-Specific Response Types in A Global Gridded Crop Model Ensemble

Crop models are often employed to project crop yields under changing conditions such as global warming and associated management change for adaptation. Multi-model ensembles are promoted to enhance the robustness of projections, but questions remain on what causes often large differences between projections of individual models. Global Gridded Crop Models (GGCMs) are especially exposed to this question when applied for assessing climate change impacts, adaptation, environmental impacts of agricultural production, because their results are used in downstream analyses, such as in integrated assessment or economic modeling for projecting future land-use change. Even though global gridded crop models are often based on detailed field-scale models or have implemented similar modeling principles in other ecosystem models, global-scale models are subject to substantial uncertainties from both model structure and parametrization as well as from calibration and input data quality. AgMIP’s Global Gridded Crop Model Intercomparison (GGCMI) has thus set out to intercompare GGCMs in order to evaluate model performance, describe model uncertainties, identify inconsistencies within the ensemble and underlying reasons, and to ultimately improve models and modeling capacities. In phase 2 of the GGCMI activities, 12 modeling groups followed a modeling protocol that asked for up to 1404 31-year global simulations at 0.5 arc-degree spatial resolution to assess models’ sensitivities to changes in carbon dioxide (C; 4 different levels) temperature (T; 7 different offset levels), water supply (W; 9 levels), and nitrogen (N; 3 levels), the so-called CTWN experiment (Franke et al. 2020; http://dx.doi.org/10.5194/gmd-13-2315-2020). We here present analyses of model response types using impact response surfaces along the C, T, W, and N dimensions, respectively and collectively. Doing so, we can understand differences in simulated responses per driver rather than aggregated changes in yields. We find that models’ sensitivities to the individual driver dimensions are substantially different and often more different across models than across regions. A cluster analysis finds regional and model-specific patterns. There is some agreement across models with respect to the spatial patterns of response types but strong differences in the distribution of response type clusters across models suggests that models need to undergo further scrutiny. We suggest establishing standards in model process evaluation not only against historical dynamics but also against dedicated experiments across the CTWN dimensions.

crop models↗

Control of the NASA Langley 16-Foot Transonic Tunnel with the Self-Organizing Feature Map

A predictive, multiple model control strategy is developed based on an ensemble of local linear models of the nonlinear system dynamics for a transonic wind tunnel. The local linear models are estimated directly from the weights of a Self Organizing Feature Map (SOFM). Local linear modeling of nonlinear autonomous systems with the SOFM is extended to a control framework where the modeled system is nonautonomous, driven by an exogenous input. This extension to a control framework is based on the consideration of a finite number of subregions in the control space. Multiple self organizing feature maps collectively model the global response of the wind tunnel to a finite set of representative prototype controls. These prototype controls partition the control space and incorporate experimental knowledge gained from decades of operation. Each SOFM models the combination of the tunnel with one of the representative controls, over the entire range of operation. The SOFM based linear models are used to predict the tunnel response to a larger family of control sequences which are clustered on the representative prototypes. The control sequence which corresponds to the prediction that best satisfies the requirements on the system output is applied as the external driving signal. Each SOFM provides a codebook representation of the tunnel dynamics corresponding to a prototype control. Different dynamic regimes are organized into topological neighborhoods where the adjacent entries in the codebook represent the minimization of a similarity metric which is the essence of the self organizing feature of the map. Thus, the SOFM is additionally employed to identify the local dynamical regime, and consequently implements a switching scheme than selects the best available model for the applied control. Experimental results of controlling the wind tunnel, with the proposed method, during operational runs where strict research requirements on the control of the Mach number were met, are presented. Comparison to similar runs under the same conditions with the tunnel controlled by either the existing controller or an expert operator indicate the superiority of the method.

Motter, Mark A.↗

The evolution of voids in the adhesion approximation

We apply the adhesion approximation to study the formation and evolution of voids in the universe. Our simulations-carried out using 128(exp 3) particles in a cubical box with side 128 Mpc-indicate that the void spectrum evolves with time and that the mean void size in the standard Cosmic Background Explorer Satellite (COBE)-normalized cold dark matter (CDM) model with H(sub 50) = 1 scals approximately as bar D(z) = bar D(sub zero)/(1+2)(exp 1/2), where bar D(sub zero) approximately = 10.5 Mpc. Interestingly, we find a strong correlation between the sizes of voids and the value of the primordial gravitational potential at void centers. This observation could in principle, pave the way toward reconstructing the form of the primordialpotential from a knowledge of the observed void spectrum. Studying the void spectrum at different cosmological epochs, for spectra with a built in k-space cutoff we find that the number of voids in a representative volume evolves with time. The mean number of voids first increases until a maximum value is reached (indicating that the formation of cellular structure is complete), and then begins to decrease as clumps and filaments erge leading to hierarchical clustering and the subsequent elimination of small voids. The cosmological epoch characterizing the completion of cellular structure occurs when the length scale going nonlinear approaches the mean distance between peaks of the gravitaional potential. A central result of this paper is that voids can be populated by substructure such as mini-sheets and filaments, which run through voids. The number of such mini-pancakes that pass through a given void can be measured by the genus characteristic of an individual void which is an indicator of the topology of a given void in intial (Lagrangian) space. Large voids have on an average a larger measure than smaller voids indicating more substructure within larger voids relative to smaller ones. We find that the topology of individual voids is strongly epoch dependent, with void topologies generally simplifying with time. This means that as voids grow older they become progressively more empty and have less structure within them. We evaluate the genus measure both for individual voids as well as for the entire ensemble of voids predicted by CDM model. As a result we find that the topology of voids when taken together with the void spectrum is a very useful statistical indicator of the evolution of the structure of the universe on large scales.

Sahni, Varun↗

Mars Ground Level Enhancements in the Context of the Solar Energetic Particle Clock

In this work we discuss the growing ensemble of solar particle events registered on the Martian surface, including their temporal appearance and solar sources. Solar energetic particle events from the surface of Mars have been observed starting soon after the August 2012 landing of the Radiation Assessment Detector onboard Curiosity. The Martian atmosphere prevents protons and heavy ions of up to 180 MeV/n kinetic energy from directly reaching the Martian surface. This cut-off is high enough to limit the number of solar energetic particle events measured on the surface to only 15 in ~12 ½ years. Yet we find in this analysis that Mars ground level enhancements follow the distinct SEP clock pattern as proton events observed at lower energies as reported in Posner, Richardson and Strauss (2024). Proton acceleration occurs predominantly near the solar surface, while transport to Mars incurs a delay in onset, and, as we show here, peak, that is a function of the longitudinal magnetic connection distance. between the foot point of solar wind magnetic field lines that intersect the Mars environment with the source longitude of the solar magnetic eruption. A distinct clustering of relative solar source locations at or near the Mars foot points at the Sun’s western limb is apparent, indicating lower flux thresholds from such preferred locations. Our findings have implications for astronaut safety at Mars.

Mars Ground Level Enhancements↗

A Taxonomy-Based Approach to Shed Light on the Babel of Mathematical Models for Rice Simulation

For most biophysical domains, differences in model structures are seldom quantified. Here, we used a taxonomy-based approach to characterise thirteen rice models. Classification keys and binary attributes for each key were identified, and models were categorised into five clusters using a binary similarity measure and the unweighted pair-group method with arithmetic mean. Principal component analysis was performed on model outputs at four sites. Results indicated that (i) differences in structure often resulted in similar predictions and (ii) similar structures can lead to large differences in model outputs. User subjectivity during calibration may have hidden expected relationships between model structure and behaviour. This explanation, if confirmed, highlights the need for shared protocols to reduce the degrees of freedom during calibration, and to limit, in turn, the risk that user subjectivity influences model performance.

model parameterisation↗

Clouds and Convective Self-Aggregation in a Multi-Model Ensemble of Radiative-Convective Equilibrium Simulations

The Radiative-Convective Equilibrium Model Intercomparison Project (RCEMIP) is an intercomparison of multiple types of numerical models configured in radiative-convective56equilibrium (RCE). RCE is an idealization of the tropical atmosphere that has long been used to study basic questions in climate science. Here, we employ RCE to investigate the role that clouds and convective activity play in determining cloud feedbacks, climatecsensitivity, the state of convective aggregation, and the equilibrium climate. RCEMIP is unique amongst intercomparisons in its inclusion of a wide range of model types, including atmospheric general circulation models (GCMs), single column models (SCMs), cloud-resolving models (CRMs), large eddy simulations (LES), and global cloud-resolving models (GCRMs). The first results are presented from the RCEMIP ensemble of more than 30 models. While there are large differences across the RCEMIP ensemble in the representation of mean profiles of temperature, humidity, and cloudiness, in a majority of models anvil clouds rise, warm, and decrease in area coverage in response to an increase in sea surface temperature (SST). Nearly all models exhibit self-aggregation in large domains and agree that self-aggregation acts to dry and warm the troposphere, reduce high cloudiness, and increase cooling to space. The degree of self-aggregation exhibits no clear tendency with warming. There is a wide range of climate sensitivities, but models with parameterized convection tend to have lower climate sensitivities than models with explicit convection. In models with parameterized convection, aggregated simulations have lower climate sensitivities than un-aggregated simulations. Plain Language Summary This study investigates tropical clouds and climate using results from more than 30 different numerical models set up in a simplified framework. The dataset of model simulations is unique in that it includes a wide range of model types configured in a consistent manner. We address some of the biggest open questions in climate science, including how cloud properties change with warming and the role that the tendency of clouds to form clusters plays in determining the average climate and how climate changes. While there are large differences in how the different models simulate average temperature, humidity, and cloudiness, in a majority of models, the amount of high clouds decreases as climate warms. Nearly all models simulate a tendency for clouds to cluster together. There is agreement that when the clouds are clustered, the atmosphere is drier with fewer clouds overall. We don’t find a conclusive result for how cloud clustering changes as the climate warms.

Allison A. Wing↗

Simulation of Radiation-Induced DNA Damage With the Code RITRACKS

INTRODUCTION DNA damage is one of the most physiologically important effects of ionizing radiation. Clustered DNA damage events, like double-strand breaks (DSBs), have the most notable biological consequences. DNA damage types depend on both the track structure of the radiation and the spatial organization of the DNA. High linear energy transfer (LET) charged nuclei, found in galactic cosmic rays (GCR), are known to produce large numbers of complex DNA damage events. The human genome is packaged into chromatin, which can take on locus-dependent and cell type-dependent spatial conformations that correspond to epigenetic states, such as more open, extended structures in transcriptionally active chromatin. These epigenetic differences can affect DNA break patterns in response to ionizing radiation, potentially creating distinct DNA repair and signaling outcomes across the genome in different cells. MATERIAL AND METHODS The code RITRACKS (Relativistic Ion Tracks), which simulates stochastic radiation track structures and radiation chemistry, was used to model damage on isolated and histone-bound DNA by various types of ions and photons. The changes made to the code to perform radiation-induced DNA damage, and simulation results on single nucleosomes are given in our recent paper. In this work, the DNA building capabilities of RITRACKS have been extended to simulate more complex DNA structures build on the coarse-grain simulation framework meso-WLCsim. This code can sample generic chromatin fiber conformation ensembles based on the geometry of nucleosomes and mechanical properties of DNA. Using RITRACKS, we simulated the fragment length distributions (FLD) of irradiated DNA structures built using the chromatin conformations of WLCsim and obtained results representative of those obtained with Radiation-Induced Correlated Cleavage with sequencing (RICC-Seq) experiments [6]. We have also performed Fe ion and photon irradiations of K562, IMR90, BJ and RPE-1 cells at NSRL to experimentally validate results. Sample processing and data analysis are in progress and any available preliminary results will be discussed. DISCUSSION The recent updates in the code RITRACKS allow the calculation of several quantities such as the DNA damage yield and the FLD. This approach can be used to model epigenetic state-specific chromatin structure parameters to leverage the epigenetic state data available for many human cell types to infer relative DNA damage sensitivity among genomic loci.

I Plante↗

Automated Knowledge Discovery From Simulators

A computational method, SimLearn, has been devised to facilitate efficient knowledge discovery from simulators. Simulators are complex computer programs used in science and engineering to model diverse phenomena such as fluid flow, gravitational interactions, coupled mechanical systems, and nuclear, chemical, and biological processes. SimLearn uses active-learning techniques to efficiently address the "landscape characterization problem." In particular, SimLearn tries to determine which regions in "input space" lead to a given output from the simulator, where "input space" refers to an abstraction of all the variables going into the simulator, e.g., initial conditions, parameters, and interaction equations. Landscape characterization can be viewed as an attempt to invert the forward mapping of the simulator and recover the inputs that produce a particular output. Given that a single simulation run can take days or weeks to complete even on a large computing cluster, SimLearn attempts to reduce costs by reducing the number of simulations needed to effect discoveries. Unlike conventional data-mining methods that are applied to static predefined datasets, SimLearn involves an iterative process in which a most informative dataset is constructed dynamically by using the simulator as an oracle. On each iteration, the algorithm models the knowledge it has gained through previous simulation trials and then chooses which simulation trials to run next. Running these trials through the simulator produces new data in the form of input-output pairs. The overall process is embodied in an algorithm that combines support vector machines (SVMs) with active learning. SVMs use learning from examples (the examples are the input-output pairs generated by running the simulator) and a principle called maximum margin to derive predictors that generalize well to new inputs. In SimLearn, the SVM plays the role of modeling the knowledge that has been gained through previous simulation trials. Active learning is used to determine which new input points would be most informative if their output were known. The selected input points are run through the simulator to generate new information that can be used to refine the SVM. The process is then repeated. SimLearn carefully balances exploration (semi-randomly searching around the input space) versus exploitation (using the current state of knowledge to conduct a tightly focused search). During each iteration, SimLearn uses not one, but an ensemble of SVMs. Each SVM in the ensemble is characterized by different hyper-parameters that control various aspects of the learned predictor - for example, whether the predictor is constrained to be very smooth (nearby points in input space lead to similar output predictions) or whether the predictor is allowed to be "bumpy." The various SVMs will have different preferences about which input points they would like to run through the simulator next. SimLearn includes a formal mechanism for balancing the ensemble SVM preferences so that a single choice can be made for the next set of trials.

Burl, Michael↗