UK researchers are teaching computers to see and label galaxies using unsupervised machine learning. The group from the University of Hertfordshire, Hatfield, presents a novel unsupervised learning approach to automatically segment and label images in astronomical surveys.
Automation of this procedure will be essential as next-generation surveys enter the petabyte scale: data volumes will exceed the capability of even large crowd-sourced analyses. We demonstrate how a growing neural gas (GNG) can be used to encode the feature space of imaging data. When coupled with a technique called hierarchical clustering, imaging data can be automatically segmented and labelled by organizing nodes in the GNG. The key distinction of unsupervised learning is that these labels need not be known prior to training, rather they are determined by the algorithm itself. Importantly, after training a network can be be presented with images it has never ‘seen’ before and provide consistent categorization of features.