Searcharxiv⌕ Search

arXiv subjects

I. Lecoeur-Taïbi

Publications and source records attributed to I. Lecoeur-Taïbi.

11 recordsLinked to original sources

Gaia Data Release 3: Gaia scan-angle dependent signals and spurious periods

Context: Gaia DR3 time series data may contain spurious signals related to the time-dependent scan angle. Aims: We aim to explain the origin of scan-angle dependent signals and how they can lead to spurious periods, provide statistics to identify them in the data, and suggest how to deal with them in Gaia DR3 data and in future releases. Methods: Using real Gaia data, alongside numerical and analytical models, we visualise and explain the features observed in the data. Results: We demonstrated with Gaia data that source structure (multiplicity or extendedness) or pollution from close-by bright objects can cause biases in the image parameter determination from which photometric, astrometric and (indirectly) radial velocity time series are derived. These biases are a function of the time-dependent scan direction of the instrument and thus can introduce scan-angle dependent signals, which in turn can result in specific spurious periodic signals. Numerical simulations qualitatively reproduce the general structure observed in the spurious period and spatial distribution of photometry and astrometry. A variety of statistics allows for identification of affected sources. Conclusions: The origin of the scan-angle dependent signals and subsequent spurious periods is well-understood and is in majority caused by fixed-orientation optical pairs with separation <0.5" (amongst which binaries with P>>5y) and (cores of) distant galaxies. Though the majority of sources with affected derived parameters have been filtered out from the Gaia archive, there remain Gaia DR3 data that should be treated with care (e.g. gaia_source was untouched). Finally, the various statistics discussed in the paper can not only be used to identify and filter affected sources, but alternatively reveal new information about them not available through other means, especially in terms of binarity on sub-arcsecond scale.

astro-ph.IM↗

Gaia Data Release 3: Cross-match of Gaia sources with variable objects from the literature

Context. In the current ever increasing data volumes of astronomical surveys, automated methods are essential. Objects of known classes from the literature are necessary for training supervised machine learning algorithms, as well as for verification/validation of their results. Aims.The primary goal of this work is to provide a comprehensive data set of known variable objects from the literature cross-matched with \textit{Gaia}~DR3 sources, including a large number of both variability types and representatives, in order to cover as much as possible sky regions and magnitude ranges relevant to each class. In addition, non-variable objects from selected surveys are targeted to probe their variability in \textit{Gaia} and possible use as standards. This data set can be the base for a training set applicable in variability detection, classification, and validation. MethodsA statistical method that employed both astrometry (position and proper motion) and photometry (mean magnitude) was applied to selected literature catalogues in order to identify the correct counterparts of the known objects in the \textit{Gaia} data. The cross-match strategy was adapted to the properties of each catalogue and the verification of results excluded dubious matches. Results.Our catalogue gathers 7\,841\,723 \textit{Gaia} sources among which 1.2~million non-variable objects and 1.7~million galaxies, in addition to 4.9~million variable sources representing over 100~variability (sub)types. Conclusions.This data set served the requirements of \textit{Gaia}'s variability pipeline for its third data release (DR3), from classifier training to result validation, and it is expected to be a useful resource for the scientific community that is interested in the analysis of variability in the \textit{Gaia} data and other surveys.

astro-ph.IM↗

Gaia Data Release 2: All-sky classification of high-amplitude pulsating stars

More than half a million of the 1.69 billion sources in Gaia Data Release 2 (DR2) are published with photometric time series that exhibit light variations during the 22 months of observation. An all-sky classification of common high-amplitude pulsators (Cepheids, long-period variables, Delta Scuti / SX Phoenicis, and RR Lyrae stars) is provided for stars with brightness variations greater than 0.1 mag in G band. A semi-supervised classification approach was employed, firstly training multi-stage random forest classifiers with sources of known types in the literature, followed by a preliminary classification of the Gaia data and a second training phase that included a selection of the first classification results to improve the representation of some classes, before the improved classifiers were applied to the Gaia data. Dedicated validation classifiers were used to reduce the level of contamination in the published results. A relevant fraction of objects were not yet sufficiently sampled for reliable Fourier series decomposition, consequently classifiers were based on features derived from statistics of photometric time series in the G, BP, and RP bands, as well as from some astrometric parameters. The published classification results include 195,780 RR Lyrae stars, 150,757 long-period variables, 8550 Cepheids, and 8882 Delta Scuti / SX Phoenicis stars. All of these results represent candidates whose completeness and contamination are described as a function of variability type and classification reliability. Results are expressed in terms of class labels and classification scores, which are available in the vari_classifier_result table of the Gaia archive.

astro-ph.SR↗

Gaia Data Release 2. Short-timescale variability processing and analysis

The Gaia DR2 sample of short-timescale variable candidates results from the investigation of the first 22 months of Gaia photometry for a subsample of sources at the Gaia faint end. For this exercise, we limited ourselves to the case of suspected rapid periodic variability. Our study combines fast-variability detection through variogram analysis, high-frequency search by means of least-squares periodograms, and empirical selection based on the investigation of specific sources seen through the Gaia eyes (e.g. known variables or visually identified objects with peculiar features in their light curves). The progressive definition and validation of this selection criterion also benefited from supplementary ground-based photometric monitoring of a few preliminary candidates, performed at the Flemish Mercator telescope (Canary Islands, Spain) between August and November 2017. We publish a list of 3,018 short-timescale variable candidates, spread throughout the sky, with a false-positive rate up to 10-20% in the Magellanic Clouds, and a more significant but justifiable contamination from longer-period variables between 19% and 50%, depending on the area of the sky. Although its completeness is limited to about 0.05%, this first sample of Gaia short-timescale variables recovers some very interesting known short-period variables, such as post-common envelope binaries or cataclysmic variables, and brings to light some fascinating, newly discovered variable sources. In the perspective of future Gaia data releases, several improvements of the short-timescale variability processing are considered, by enhancing the existing variogram and period-search algorithms or by classifying the identified candidates. Nonetheless, the encouraging outcome of our Gaia DR2 analysis demonstrates the power of this mission for such fast-variability studies, and opens great perspectives for this domain of astrophysics.

astro-ph.IM↗

Gaia Data Release 2: The first Gaia catalogue of long-period variable candidates

Gaia DR2 provides a unique all-sky catalogue of 550'737 variable stars, of which 151'761 are long-period variable (LPV) candidates with G variability amplitudes larger than 0.2 mag (5-95% quantile range). About one-fifth of the LPV candidates are Mira candidates, the majority of the rest are semi-regular variable candidates. For each source, G, BP , and RP photometric time-series are published, together with some LPV-specific attributes for the subset of 89'617 candidates with periods in G longer than 60 days. We describe this first Gaia catalogue of LPV candidates, and present various validation checks. Various samples of LPVs were used to validate the catalogue: a sample of well-studied very bright LPVs with light curves from the AAVSO that are partly contemporaneous with Gaia light curves, a sample of Gaia LPV candidates with good parallaxes, the ASAS_SN catalogue of LPVs, and the OGLE catalogues of LPVs towards the Magellanic Clouds and the Galactic bulge. The analyses of these samples show a good agreement between Gaia DR2 and literature periods. The same is globally true for bolometric corrections of M-type stars. The main contaminant of our DR2 catalogue comes from young stellar objects (YSOs) in the solar vicinity (within ~1 kpc), although their number in the whole catalogue is only at the percent level. A cautionary note is provided about parallax-dependent LPV attributes published in the catalogue. This first Gaia catalogue of LPVs approximately doubles the number of known LPVs with amplitudes larger than 0.2 mag, despite the conservative candidate selection criteria that prioritise low contamination over high completeness, and despite the limited DR2 time coverage compared to the long periods characteristic of LPVs. It also contains a small set of YSO candidates, which offers the serendipitous opportunity to study these objects at an early stage of the Gaia data releases.

astro-ph.SR↗

Gaia Data Release 2: Summary of the variability processing & analysis results

The Gaia Data Release 2 (DR2): we summarise the processing and results of the identification of variable source candidates of RR Lyrae stars, Cepheids, long period variables (LPVs), rotation modulation (BY Dra-type) stars, delta Scuti & SX Phoenicis stars, and short-timescale variables. In this release we aim to provide useful but not necessarily complete samples of candidates. The processed Gaia data consist of the G, BP, and RP photometry during the first 22 months of operations as well as positions and parallaxes. Various methods from classical statistics, data mining and time series analysis were applied and tailored to the specific properties of Gaia data, as well as various visualisation tools. The DR2 variability release contains: 228'904 RR Lyrae stars, 11'438 Cepheids, 151'761 LPVs, 147'535 stars with rotation modulation, 8'882 delta Scuti & SX Phoenicis stars, and 3'018 short-timescale variables. These results are distributed over a classification and various Specific Object Studies (SOS) tables in the Gaia archive, along with the three-band time series and associated statistics for the underlying 550'737 unique sources. We estimate that about half of them are newly identified variables. The variability type completeness varies strongly as function of sky position due to the non-uniform sky coverage and intermediate calibration level of this data. The probabilistic and automated nature of this work implies certain completeness and contamination rates which are quantified so that users can anticipate their effects. This means that even well-known variable sources can be missed or misidentified in the published data. The DR2 variability release only represents a small subset of the processed data. Future releases will include more variable sources and data products; however, DR2 shows the (already) very high quality of the data and great promise for variability studies.

astro-ph.SR↗

Gaia eclipsing binary and multiple systems. Two-Gaussian models applied to OGLE-III eclipsing binary light curves in the Large Magellanic Cloud

The advent of large scale multi-epoch surveys raises the need for automated light curve (LC) processing. This is particularly true for eclipsing binaries (EBs), which form one of the most populated types of variable objects. The Gaia mission, launched at the end of 2013, is expected to detect of the order of few million EBs over a 5-year mission. We present an automated procedure to characterize EBs based on the geometric morphology of their LCs with two aims: first to study an ensemble of EBs on a statistical ground without the need to model the binary system, and second to enable the automated identification of EBs that display atypical LCs. We model the folded LC geometry of EBs using up to two Gaussian functions for the eclipses and a cosine function for any ellipsoidal-like variability that may be present between the eclipses. The procedure is applied to the OGLE-III data set of EBs in the Large Magellanic Cloud (LMC) as a proof of concept. The bayesian information criterion is used to select the best model among models containing various combinations of those components, as well as to estimate the significance of the components. Based on the two-Gaussian models, EBs with atypical LC geometries are successfully identified in two diagrams, using the Abbe values of the original and residual folded LCs, and the reduced $χ^2$. Cleaning the data set from the atypical cases and further filtering out LCs that contain non-significant eclipse candidates, the ensemble of EBs can be studied on a statistical ground using the two-Gaussian model parameters. For illustration purposes, we present the distribution of projected eccentricities as a function of orbital period for the OGLE-III set of EBs in the LMC, as well as the distribution of their primary versus secondary eclipse widths.

astro-ph.IM↗

Gaia eclipsing binary and multiple systems. Supervised classification and self-organizing maps

Large surveys producing tera- and petabyte-scale databases require machine-learning and knowledge discovery methods to deal with the overwhelming quantity of data and the difficulties of extracting concise, meaningful information with reliable assessment of its uncertainty. This study investigates the potential of a few machine-learning methods for the automated analysis of eclipsing binaries in the data of such surveys. We aim to aid the extraction of samples of eclipsing binaries from such databases and to provide basic information about the objects. We estimate class labels according to two classification systems, one based on the light curve morphology (EA/EB/EW classes) and the other based on the physical characteristics of the binary system (system morphology classes; detached through overcontact systems). Furthermore, we explore low-dimensional surfaces along which the light curves of eclipsing binaries are concentrated, to use in the characterization of the binary systems and in the exploration of biases of the full unknown Gaia data with respect to the training sets. We explore the performance of principal component analysis (PCA), linear discriminant analysis (LDA), random forest classification and self-organizing maps (SOM). We pre-process the photometric time series by combining a double Gaussian profile fit and a smoothing spline, in order to de-noise and interpolate the observed light curves. We achieve further denoising, and selected the most important variability elements from the light curves using PCA. We perform supervised classification using random forest and LDA based on the PC decomposition, while SOM gives a continuous 2-dimensional manifold of the light curves arranged by a few important features. We estimate the uncertainty of the supervised methods due to the specific finite training set using ensembles of models constructed on randomized training sets.

astro-ph.IM↗

Automated classification of Hipparcos unsolved variables

We present an automated classification of stars exhibiting periodic, non-periodic and irregular light variations. The Hipparcos catalogue of unsolved variables is employed to complement the training set of periodic variables of Dubath et al. with irregular and non-periodic representatives, leading to 3881 sources in total which describe 24 variability types. The attributes employed to characterize light-curve features are selected according to their relevance for classification. Classifier models are produced with random forests and a multistage methodology based on Bayesian networks, achieving overall misclassification rates under 12 per cent. Both classifiers are applied to predict variability types for 6051 Hipparcos variables associated with uncertain or missing types in the literature.

astro-ph.SR↗

Long Period Variables in the Large Magellanic Cloud from the EROS-2 survey

Context. The EROS-2 survey has produced a database of millions of time series from stars monitored for more than six years, allowing to classify some of their sources into different variable star types. Among these, Long Period Variables (LPVs), known to follow sequences in the period-luminosity diagram, include long secondary period variables whose variability origin is still a matter of debate. Aims.We use the 856 864 variable stars available from the Large Magellanic Cloud (LMC) in the EROS-2 database to detect, classify and characterize LPVs. Methods. Our method to extract LPVs is based on the statistical Abbe test. It investigates the regularity of the light curve with respect to the survey duration in order to extract candidates with long-term variability. The period search is done by Deeming, Lomb-Scargle and generalized Lomb-Scargle methods, combined with Fourier series fit. Color-magnitude, period-magnitude and period-amplitude diagrams are used to characterize our candidates. Results. We present a catalog of 43 551 LPV candidates for the Large Magellanic Cloud. For each of them, we provide up to five periods, mean magnitude in EROS-2, 2MASS and Spitzer bands, BE-RE color, RE amplitude and spectral type.We use infrared data to make the distinction between RGB, O-rich, C-rich and extreme AGB stars. Properties of our LPV candidates are investigated by analyzing period-luminosity and period-amplitude diagrams.

astro-ph.SR↗

Hipparcos Variable Star Detection and Classification Efficiency

A complete periodic star extraction and classification scheme is set up and tested with the Hipparcos catalogue. The efficiency of each step is derived by comparing the results with prior knowledge coming from the catalogue or from the literature. A combination of two variability criteria is applied in the first step to select 17 006 variability candidates from a complete sample of 115 152 stars. Our candidate sample turns out to include 10 406 known variables (i.e., 90% of the total of 11 597) and 6600 contaminating constant stars. A random forest classification is used in the second step to extract 1881 (82%) of the known periodic objects while removing entirely constant stars from the sample and limiting the contamination of non-periodic variables to 152 stars (7.5%). The confusion introduced by these 152 non-periodic variables is evaluated in the third step using the results of the Hipparcos periodic star classification presented in a previous study (Dubath et al. [1]).

astro-ph.SR↗