Engineering PapersSearch

DOE OSTI · 3367298

Cohort organized learning: clustering through agreement

Abstract

In this article we describe cohort organized learning (CoOL), a method for clustering data without explicit distance or similarity computations. Herein, we will describe CoOL, derive the gradients determined by expectation maximization to train the networks, show how to monitor convergence during training and evaluate the clusters after training, and discuss a series of examples and use cases. We also discuss CoOL’s limitations and future prospects on related tasks. Because CoOL uses neural networks to estimate the clusters, it can be used to cluster any data that can be made compatible and we illustrate this on vector data and images.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

O’Shea, Finn H. [SLAC National Accelerator Laboratory (SLAC), Menlo Park, CA (United States)] (ORCID:0000000323987381), Elena Monzani, Maria [SLAC National Accelerator Laboratory (SLAC), Menlo Park, CA (United States); Stanford Univ., CA (United States). Kavli Institute for Particle Astrophysics & Cosmology] (ORCID:0000000282545308). 2026-06-16. Cohort organized learning: clustering through agreement. https://doi.org/10.1088/2632-2153%2Fae779f

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related reports

Concept Lens: Visual Comparison and Evaluation of Generative Model Manipulations

Generative models are becoming a transformative technology for the creation and editing of images. However, it remains challenging to harness these models for precise image manipulation. These challenges often manifest as inconsistency in the editing process, where both the type and amount of semantic change, depend on the image being manipulated. Moreover, there exist many methods for computing image manipulations, whose development is hindered by the matter of inconsistency. This paper aims to address these challenges by improving how we evaluate, compare, and explore the space of manipulations offered by a generative model. We present Concept Lens, a visual interface that is designed to aid users in understanding semantic concepts carried in image manipulations, and how these manipulations vary over generated images. Given the large space of possible images produced by a generative model, Concept Lens is designed to support the exploration of both generated images, and their manipulations, at multiple levels of detail. To this end, the layout of Concept Lens is informed by two hierarchies: a hierarchical organization of (1) original images, grouped by their similarities, and (2) image manipulations, where manipulations that induce similar changes are grouped together. This layout allows one to discover the types of images that consistently respond to a group of manipulations, and vice versa, manipulations that consistently respond to a group of codes. We show the benefits of this design across multiple use cases, specifically, studying the quality of manipulations for a single method, and offering a means of comparing different methods.

clustering

A Deterministic Annealing Approach to Clustering AIRS Data

We will examine the validity of means and standard deviations as a basis for climate data products. We will explore the conditions under which these two simple statistics are inadequate summaries of the underlying empirical probability distributions by contrasting them with a nonparametric, method called Deterministic Annealing technique

clustering