Gaia eclipsing binary and multiple systems. Supervised classification and self-organizing maps

Süveges, M.; Barblan, F.; Lecoeur-Taïbi, I.; Prša, A.; Holl, B.; Eyer, L.; Kochoska, A.; Mowlavi, N.; Rimoldini, L.

doi:10.1051/0004-6361/201629710

Astrophysics > Instrumentation and Methods for Astrophysics

arXiv:1702.06296 (astro-ph)

[Submitted on 21 Feb 2017]

Title:Gaia eclipsing binary and multiple systems. Supervised classification and self-organizing maps

Authors:M. Süveges, F. Barblan, I. Lecoeur-Taïbi, A. Prša, B. Holl, L. Eyer, A. Kochoska, N. Mowlavi, L. Rimoldini

View PDF

Abstract:Large surveys producing tera- and petabyte-scale databases require machine-learning and knowledge discovery methods to deal with the overwhelming quantity of data and the difficulties of extracting concise, meaningful information with reliable assessment of its uncertainty. This study investigates the potential of a few machine-learning methods for the automated analysis of eclipsing binaries in the data of such surveys. We aim to aid the extraction of samples of eclipsing binaries from such databases and to provide basic information about the objects. We estimate class labels according to two classification systems, one based on the light curve morphology (EA/EB/EW classes) and the other based on the physical characteristics of the binary system (system morphology classes; detached through overcontact systems). Furthermore, we explore low-dimensional surfaces along which the light curves of eclipsing binaries are concentrated, to use in the characterization of the binary systems and in the exploration of biases of the full unknown Gaia data with respect to the training sets. We explore the performance of principal component analysis (PCA), linear discriminant analysis (LDA), random forest classification and self-organizing maps (SOM). We pre-process the photometric time series by combining a double Gaussian profile fit and a smoothing spline, in order to de-noise and interpolate the observed light curves. We achieve further denoising, and selected the most important variability elements from the light curves using PCA. We perform supervised classification using random forest and LDA based on the PC decomposition, while SOM gives a continuous 2-dimensional manifold of the light curves arranged by a few important features. We estimate the uncertainty of the supervised methods due to the specific finite training set using ensembles of models constructed on randomized training sets.

Comments:	20 pages, 22 figures. Accepted for publication in A&A
Subjects:	Instrumentation and Methods for Astrophysics (astro-ph.IM); Solar and Stellar Astrophysics (astro-ph.SR)
Cite as:	arXiv:1702.06296 [astro-ph.IM]
	(or arXiv:1702.06296v1 [astro-ph.IM] for this version)
	https://doi.org/10.48550/arXiv.1702.06296
Journal reference:	A&A 603, A117 (2017)
Related DOI:	https://doi.org/10.1051/0004-6361/201629710

Submission history

From: Maria Süveges Dr [view email]
[v1] Tue, 21 Feb 2017 09:04:02 UTC (515 KB)

Astrophysics > Instrumentation and Methods for Astrophysics

Title:Gaia eclipsing binary and multiple systems. Supervised classification and self-organizing maps

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Astrophysics > Instrumentation and Methods for Astrophysics

Title:Gaia eclipsing binary and multiple systems. Supervised classification and self-organizing maps

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators