Semi-supervised classification and clustering analysis for variable stars. Issue 3 (30th September 2022)
- Record Type:
- Journal Article
- Title:
- Semi-supervised classification and clustering analysis for variable stars. Issue 3 (30th September 2022)
- Main Title:
- Semi-supervised classification and clustering analysis for variable stars
- Authors:
- Pantoja, R
Catelan, M
Pichara, K
Protopapas, P - Abstract:
- ABSTRACT: The immense amount of time series data produced by astronomical surveys has called for the use of machine learning algorithms to discover and classify several million celestial sources. In the case of variable stars, supervised learning approaches have become commonplace. However, this needs a considerable collection of expert-labelled light curves to achieve adequate performance, which is costly to construct. To solve this problem, we introduce two approaches. First, a semi-supervised hierarchical method, which requires substantially less trained data than supervised methods. Second, a clustering analysis procedure that finds groups that may correspond to classes or subclasses of variable stars. Both methods are primarily supported by dimensionality reduction of the data for visualization and to avoid the curse of dimensionality. We tested our methods with catalogues collected from the Optical Gravitational Lensing Experiment (OGLE), the Catalina Sky Survey (CSS), and the Gaia survey. The semi-supervised method reaches a performance of around 90 per cent for all of our three selected catalogues of variable stars using only $5{{\ \rm per\ cent}}$ of the data in the training. This method is suitable for classifying the main classes of variable stars when there is only a small amount of training data. Our clustering analysis confirms that most of the clusters found have a purity over 90 per cent with respect to classes and 80 per cent with respect to subclasses,ABSTRACT: The immense amount of time series data produced by astronomical surveys has called for the use of machine learning algorithms to discover and classify several million celestial sources. In the case of variable stars, supervised learning approaches have become commonplace. However, this needs a considerable collection of expert-labelled light curves to achieve adequate performance, which is costly to construct. To solve this problem, we introduce two approaches. First, a semi-supervised hierarchical method, which requires substantially less trained data than supervised methods. Second, a clustering analysis procedure that finds groups that may correspond to classes or subclasses of variable stars. Both methods are primarily supported by dimensionality reduction of the data for visualization and to avoid the curse of dimensionality. We tested our methods with catalogues collected from the Optical Gravitational Lensing Experiment (OGLE), the Catalina Sky Survey (CSS), and the Gaia survey. The semi-supervised method reaches a performance of around 90 per cent for all of our three selected catalogues of variable stars using only $5{{\ \rm per\ cent}}$ of the data in the training. This method is suitable for classifying the main classes of variable stars when there is only a small amount of training data. Our clustering analysis confirms that most of the clusters found have a purity over 90 per cent with respect to classes and 80 per cent with respect to subclasses, suggesting that this type of analysis can be used in large-scale variability surveys as an initial step to identify which classes or subclasses of variable stars are present in the data and/or to build training sets, among many other possible applications. … (more)
- Is Part Of:
- Monthly notices of the Royal Astronomical Society. Volume 517:Issue 3(2022)
- Journal:
- Monthly notices of the Royal Astronomical Society
- Issue:
- Volume 517:Issue 3(2022)
- Issue Display:
- Volume 517, Issue 3 (2022)
- Year:
- 2022
- Volume:
- 517
- Issue:
- 3
- Issue Sort Value:
- 2022-0517-0003-0000
- Page Start:
- 3660
- Page End:
- 3681
- Publication Date:
- 2022-09-30
- Subjects:
- methods: data analysis -- methods: statistical -- stars: variables: general
Astronomy -- Periodicals
Periodicals
520.5 - Journal URLs:
- http://mnras.oxfordjournals.org/ ↗
http://onlinelibrary.wiley.com/journal/10.1111/(ISSN)1365-2966 ↗
http://www.blackwell-synergy.com/issuelist.asp?journal=mnr ↗
http://www.blackwell-synergy.com/loi/mnr ↗
http://ukcatalogue.oup.com/ ↗ - DOI:
- 10.1093/mnras/stac2715 ↗
- Languages:
- English
- ISSNs:
- 0035-8711
- Deposit Type:
- Legaldeposit
- View Content:
- Available online (eLD content is only available in our Reading Rooms) ↗
- Physical Locations:
- British Library DSC - 5943.000000
British Library DSC - BLDSS-3PM
British Library HMNTS - ELD Digital store - Ingest File:
- 24196.xml