SearcharxivSearch

arXiv subjects

Laura Trouille

Publications and source records attributed to Laura Trouille.

16 recordsLinked to original sources

NGTS-EB-8: A double-lined eclipsing M+M binary discovered by citizen scientists

We report the identification and characterization of a new binary system composed of two near-equal mass M-dwarfs. The binary NGTS-EB-8 was identified as a planet candidate in data from the Next Generation Transit Survey (NGTS) by citizen scientists participating in the Planet Hunters NGTS project. High-resolution spectroscopic observations reveal the system to be a double-lined binary. By modeling the photometric and radial velocity observations, we determine an orbital period of 4.2 days and the masses and radii of both stars to be $M_A=0.250^{+0.005}_{-0.004}$ M$_{\odot}$, $M_B=0.208^{+0.005}_{-0.004}$ M$_{\odot}$, $R_A=0.255^{+0.004}_{-0.005}$ R$_{\odot}$, $R_B=0.233^{+0.006}_{-0.005}$ R$_{\odot}$. We detect Balmer line emission from at least one of the stars but no significant flare activity. We note that both components lie in the fully convective regime of low-mass stars ($<0.35$ M$_{\odot}$), therefore can be a valuable test for stellar evolutionary models. We demonstrate that the photometric observations, speckle imaging and initial radial velocity measurements were unable to identify the true nature of this system and highlight that high-resolution spectroscopic observations are crucial in determining whether systems such as this are in fact binaries.

astro-ph.SR

Planet Hunters NGTS: New Planet Candidates from a Citizen Science Search of the Next Generation Transit Survey Public Data

We present the results from the first two years of the Planet Hunters NGTS citizen science project, which searches for transiting planet candidates in data from the Next Generation Transit Survey (NGTS) by enlisting the help of members of the general public. Over 8,000 registered volunteers reviewed 138,198 light curves from the NGTS Public Data Releases 1 and 2. We utilize a user weighting scheme to combine the classifications of multiple users to identify the most promising planet candidates not initially discovered by the NGTS team. We highlight the five most interesting planet candidates detected through this search, which are all candidate short-period giant planets. This includes the TIC-165227846 system that, if confirmed, would be the lowest-mass star to host a close-in giant planet. We assess the detection efficiency of the project by determining the number of confirmed planets from the NASA Exoplanet Archive and TESS Objects of Interest (TOIs) successfully recovered by this search and find that 74% of confirmed planets and 63% of TOIs detected by NGTS are recovered by the Planet Hunters NGTS project. The identification of new planet candidates shows that the citizen science approach can provide a complementary method to the detection of exoplanets with ground-based surveys such as NGTS.

astro-ph.EP

Workforce Development in Astronomy and Astroinformatics

Policy Brief on "Workforce Development in Astronomy and Astroinformatics", distilled from the corresponding panel that was part of the discussions during S20 Policy Webinar on Astroinformatics for Sustainable Development held on 6-7 July 2023. The discipline of astronomy and astroinformatics is dynamically evolving thereby creating a compelling opportunity to foster a more inclusive, diverse, and proficient workforce. This is crucial for addressing multifaceted challenges that emerge as we progress and harness the potential therein. To realize this goal, it's imperative to cultivate strategies that promote inclusive practices in STEM education, encourage participation from historically excluded groups, provide training and mentorship, as well as provide active champions, especially for students and early career professionals from (historically) excluded groups. We provide an overview of the current status, resources available, and possible steps especially keeping in mind large international projects. The policy webinar took place during the G20 presidency in India (2023). A summary based on the seven panels can be found here: arxiv:2401.04623.

astro-ph.IM

AstroInformatics: Recommendations for Global Cooperation

Policy Brief on "AstroInformatics, Recommendations for Global Collaboration", distilled from panel discussions during S20 Policy Webinar on Astroinformatics for Sustainable Development held on 6-7 July 2023. The deliberations encompassed a wide array of topics, including broad astroinformatics, sky surveys, large-scale international initiatives, global data repositories, space-related data, regional and international collaborative efforts, as well as workforce development within the field. These discussions comprehensively addressed the current status, notable achievements, and the manifold challenges that the field of astroinformatics currently confronts. The G20 nations present a unique opportunity due to their abundant human and technological capabilities, coupled with their widespread geographical representation. Leveraging these strengths, significant strides can be made in various domains. These include, but are not limited to, the advancement of STEM education and workforce development, the promotion of equitable resource utilization, and contributions to fields such as Earth Science and Climate Science. We present a concise overview, followed by specific recommendations that pertain to both ground-based and space data initiatives. Our team remains readily available to furnish further elaboration on any of these proposals as required. Furthermore, we anticipate further engagement during the upcoming G20 presidencies in Brazil (2024) and South Africa (2025) to ensure the continued discussion and realization of these objectives. The policy webinar took place during the G20 presidency in India (2023). Notes based on the seven panels will be separately published.

astro-ph.IM

TCuPGAN: A novel framework developed for optimizing human-machine interactions in citizen science

In the era of big data in scientific research, there is a necessity to leverage techniques which reduce human effort in labeling and categorizing large datasets by involving sophisticated machine tools. To combat this problem, we present a novel, general purpose model for 3D segmentation that leverages patch-wise adversariality and Long Short-Term Memory to encode sequential information. Using this model alongside citizen science projects which use 3D datasets (image cubes) on the Zooniverse platforms, we propose an iterative human-machine optimization framework where only a fraction of the 2D slices from these cubes are seen by the volunteers. We leverage the patch-wise discriminator in our model to provide an estimate of which slices within these image cubes have poorly generalized feature representations, and correspondingly poor machine performance. These images with corresponding machine proposals would be presented to volunteers on Zooniverse for correction, leading to a drastic reduction in the volunteer effort on citizen science projects. We trained our model on ~2300 liver tissue 3D electron micrographs. Lipid droplets were segmented within these images through human annotation via the `Etch A Cell - Fat Checker' citizen science project, hosted on the Zooniverse platform. In this work, we demonstrate this framework and the selection methodology which resulted in a measured reduction in volunteer effort by more than 60%. We envision this type of joint human-machine partnership will be of great use on future Zooniverse projects.

cs.HC

Gravity Spy: Lessons Learned and a Path Forward

The Gravity Spy project aims to uncover the origins of glitches, transient bursts of noise that hamper analysis of gravitational-wave data. By using both the work of citizen-science volunteers and machine-learning algorithms, the Gravity Spy project enables reliable classification of glitches. Citizen science and machine learning are intrinsically coupled within the Gravity Spy framework, with machine-learning classifications providing a rapid first-pass classification of the dataset and enabling tiered volunteer training, and volunteer-based classifications verifying the machine classifications, bolstering the machine-learning training set and identifying new morphological classes of glitches. These classifications are now routinely used in studies characterizing the performance of the LIGO gravitational-wave detectors. Providing the volunteers with a training framework that teaches them to classify a wide range of glitches, as well as additional tools to aid their investigations of interesting glitches, empowers them to make discoveries of new classes of glitches. This demonstrates that, when giving suitable support, volunteers can go beyond simple classification tasks to identify new features in data at a level comparable to domain experts. The Gravity Spy project is now providing volunteers with more complicated data that includes auxiliary monitors of the detector to identify the root cause of glitches.

gr-qc

From fat droplets to floating forests: cross-domain transfer learning using a PatchGAN-based segmentation model

Many scientific domains gather sufficient labels to train machine algorithms through human-in-the-loop techniques provided by the Zooniverse.org citizen science platform. As the range of projects, task types and data rates increase, acceleration of model training is of paramount concern to focus volunteer effort where most needed. The application of Transfer Learning (TL) between Zooniverse projects holds promise as a solution. However, understanding the effectiveness of TL approaches that pretrain on large-scale generic image sets vs. images with similar characteristics possibly from similar tasks is an open challenge. We apply a generative segmentation model on two Zooniverse project-based data sets: (1) to identify fat droplets in liver cells (FatChecker; FC) and (2) the identification of kelp beds in satellite images (Floating Forests; FF) through transfer learning from the first project. We compare and contrast its performance with a TL model based on the COCO image set, and subsequently with baseline counterparts. We find that both the FC and COCO TL models perform better than the baseline cases when using >75% of the original training sample size. The COCO-based TL model generally performs better than the FC-based one, likely due to its generalized features. Our investigations provide important insights into usage of TL approaches on multi-domain data hosted across different Zooniverse projects, enabling future projects to accelerate task completion.

cs.LG

Survey of Gravitationally-lensed Objects in HSC Imaging (SuGOHI). VI. Crowdsourced lens finding with Space Warps

Strong lenses are extremely useful probes of the distribution of matter on galaxy and cluster scales at cosmological distances, but are rare and difficult to find. The number of currently known lenses is on the order of 1,000. We wish to use crowdsourcing to carry out a lens search targeting massive galaxies selected from over 442 square degrees of photometric data from the Hyper Suprime-Cam (HSC) survey. We selected a sample of $\sim300,000$ galaxies with photometric redshifts in the range $0.2 < z_{phot} < 1.2$ and photometrically inferred stellar masses $\log{M_*} > 11.2$. We crowdsourced lens finding on this sample of galaxies on the Zooniverse platform, as part of the Space Warps project. The sample was complemented by a large set of simulated lenses and visually selected non-lenses, for training purposes. Nearly 6,000 citizen volunteers participated in the experiment. In parallel, we used YattaLens, an automated lens finding algorithm, to look for lenses in the same sample of galaxies. Based on a statistical analysis of classification data from the volunteers, we selected a sample of the most promising $\sim1,500$ candidates which we then visually inspected: half of them turned out to be possible (grade C) lenses or better. Including lenses found by YattaLens or serendipitously noticed in the discussion section of the Space Warps website, we were able to find 14 definite lenses, 129 probable lenses and 581 possible lenses. YattaLens found half the number of lenses discovered via crowdsourcing. Crowdsourcing is able to produce samples of lens candidates with high completeness and purity, compared to currently available automated algorithms. A hybrid approach, in which the visual inspection of samples of lens candidates pre-selected by discovery algorithms and/or coupled to machine learning is crowdsourced, will be a viable option for lens finding in the 2020s.

astro-ph.IM

Planet Hunters TESS II: Findings from the first two years of TESS

We present the results from the first two years of the Planet Hunters TESS citizen science project, which identifies planet candidates in the TESS data by engaging members of the general public. Over 22,000 citizen scientists from around the world visually inspected the first 26 Sectors of TESS data in order to help identify transit-like signals. We use a clustering algorithm to combine these classifications into a ranked list of events for each sector, the top 500 of which are then visually vetted by the science team. We assess the detection efficiency of this methodology by comparing our results to the list of TESS Objects of Interest (TOIs) and show that we recover 85 % of the TOIs with radii greater than 4 Earth radii and 51 % of those with radii between 3 and 4 Earth radii. Additionally, we present our 90 most promising planet candidates that had not previously been identified by other teams, 73 of which exhibit only a single transit event in the TESS light curve, and outline our efforts to follow these candidates up using ground-based observatories. Finally, we present noteworthy stellar systems that were identified through the Planet Hunters TESS project.

astro-ph.EP

Optimizing the Human-Machine Partnership with Zooniverse

Over the past decade, Citizen Science has become a proven method of distributed data analysis, enabling research teams from diverse domains to solve problems involving large quantities of data with complexity levels which require human pattern recognition capabilities. With over 120 projects built reaching nearly 1.7 million volunteers, the Zooniverse.org platform has led the way in the application of Citizen Science as a method for closing the Big Data analysis gap. Since the launch in 2007 of the Galaxy Zoo project, the Zooniverse platform has enabled significant contributions across many disciplines; e.g., in ecology, humanities, and astronomy. Citizen science as an approach to Big Data combines the twin advantages of the ability to scale analysis to the size of modern datasets with the ability of humans to make serendipitous discoveries. To cope with the larger datasets looming on the horizon such as astronomy's Large Synoptic Survey Telescope (LSST) or the 100's of TB from ecology projects annually, Zooniverse has been researching a system design that is optimized for efficiency in task assignment and incorporating human and machine classifiers into the classification engine. By making efficient use of smart task assignment and the combination of human and machine classifiers, we can achieve greater accuracy and flexibility than has been possible to date. We note that creating the most efficient system must consider how best to engage and retain volunteers as well as make the most efficient use of their classifications. Our work thus focuses on understanding the factors that optimize efficiency of the combined human-machine system. This paper summarizes some of our research to date on integration of machine learning with Zooniverse, while also describing new infrastructure developed on the Zooniverse platform to carry out this research.

cs.HC

Floating Forests: Quantitative Validation of Citizen Science Data Generated From Consensus Classifications

Large-scale research endeavors can be hindered by logistical constraints limiting the amount of available data. For example, global ecological questions require a global dataset, and traditional sampling protocols are often too inefficient for a small research team to collect an adequate amount of data. Citizen science offers an alternative by crowdsourcing data collection. Despite growing popularity, the community has been slow to embrace it largely due to concerns about quality of data collected by citizen scientists. Using the citizen science project Floating Forests (http://floatingforests.org), we show that consensus classifications made by citizen scientists produce data that is of comparable quality to expert generated classifications. Floating Forests is a web-based project in which citizen scientists view satellite photographs of coastlines and trace the borders of kelp patches. Since launch in 2014, over 7,000 citizen scientists have classified over 750,000 images of kelp forests largely in California and Tasmania. Images are classified by 15 users. We generated consensus classifications by overlaying all citizen classifications and assessed accuracy by comparing to expert classifications. Matthews correlation coefficient (MCC) was calculated for each threshold (1-15), and the threshold with the highest MCC was considered optimal. We showed that optimal user threshold was 4.2 with an MCC of 0.400 (0.023 SE) for Landsats 5 and 7, and a MCC of 0.639 (0.246 SE) for Landsat 8. These results suggest that citizen science data derived from consensus classifications are of comparable accuracy to expert classifications. Citizen science projects should implement methods such as consensus classification in conjunction with a quantitative comparison to expert generated classifications to avoid concerns about data quality.

physics.soc-ph

A transient search using combined human and machine classifications

Large modern surveys require efficient review of data in order to find transient sources such as supernovae, and to distinguish such sources from artefacts and noise. Much effort has been put into the development of automatic algorithms, but surveys still rely on human review of targets. This paper presents an integrated system for the identification of supernovae in data from Pan-STARRS1, combining classifications from volunteers participating in a citizen science project with those from a convolutional neural network. The unique aspect of this work is the deployment, in combination, of both human and machine classifications for near real-time discovery in an astronomical project. We show that the combination of the two methods outperforms either one used individually. This result has important implications for the future development of transient searches, especially in the era of LSST and other large-throughput surveys.

astro-ph.IM

The First Brown Dwarf Discovered by the Backyard Worlds: Planet 9 Citizen Science Project

The Wide-field Infrared Survey Explorer (WISE) is a powerful tool for finding nearby brown dwarfs and searching for new planets in the outer solar system, especially with the incorporation of NEOWISE and NEOWISE-Reactivation data. So far, searches for brown dwarfs in WISE data have yet to take advantage of the full depth of the WISE images. To efficiently search this unexplored space via visual inspection, we have launched a new citizen science project, called "Backyard Worlds: Planet 9," which asks volunteers to examine short animations composed of difference images constructed from time-resolved WISE coadds. We report the discovery of the first new substellar object found by this project, WISEA J110125.95+540052.8, a T5.5 brown dwarf located approximately 34 pc from the Sun with a total proper motion of $\sim$0.7 as yr$^{-1}$. WISEA J110125.95+540052.8 has a WISE $W2$ magnitude of $W2=15.37 \pm 0.09$, this discovery demonstrates the ability of citizen scientists to identify moving objects via visual inspection that are 0.9 magnitudes fainter than the $W2$ single-exposure sensitivity, a threshold that has limited prior motion-based brown dwarf searches with WISE.

astro-ph.SR

Gravity Spy: Integrating Advanced LIGO Detector Characterization, Machine Learning, and Citizen Science

(abridged for arXiv) With the first direct detection of gravitational waves, the Advanced Laser Interferometer Gravitational-wave Observatory (LIGO) has initiated a new field of astronomy by providing an alternate means of sensing the universe. The extreme sensitivity required to make such detections is achieved through exquisite isolation of all sensitive components of LIGO from non-gravitational-wave disturbances. Nonetheless, LIGO is still susceptible to a variety of instrumental and environmental sources of noise that contaminate the data. Of particular concern are noise features known as glitches, which are transient and non-Gaussian in their nature, and occur at a high enough rate so that accidental coincidence between the two LIGO detectors is non-negligible. In this paper we describe an innovative project that combines crowdsourcing with machine learning to aid in the challenging task of categorizing all of the glitches recorded by the LIGO detectors. Through the Zooniverse platform, we engage and recruit volunteers from the public to categorize images of glitches into pre-identified morphological classes and to discover new classes that appear as the detectors evolve. In addition, machine learning algorithms are used to categorize images after being trained on human-classified examples of the morphological classes. Leveraging the strengths of both classification methods, we create a combined method with the aim of improving the efficiency and accuracy of each individual classifier. The resulting classification and characterization should help LIGO scientists to identify causes of glitches and subsequently eliminate them from the data or the detector entirely, thereby improving the rate and accuracy of gravitational-wave observations. We demonstrate these methods using a small subset of data from LIGO's first observing run.

gr-qc

A Spectroscopic Survey of WISE-selected Obscured Quasars with the Southern African Large Telescope

We present the results of an optical spectroscopic survey of a sample of 40 candidate obscured quasars identified on the basis of their mid-infrared emission detected by the Wide-Field Infrared Survey Explorer (WISE). Optical spectra for this survey were obtained using the Robert Stobie Spectrograph (RSS) on the Southern African Large Telescope (SALT). Our sample was selected with WISE colors characteristic of AGNs, as well as red optical to mid-IR colors indicating that the optical/UV AGN continuum is obscured by dust. We obtain secure redshifts for the majority of the objects that comprise our sample (35/40), and find that sources that are bright in the WISE W4 (22$μ$m) band are typically at moderate redshift ( = 0.35$) while sources fainter in W4 are at higher redshifts ( = 0.73$). The majority of the sources have narrow emission lines, with optical colors and emission line ratios of our WISE-selected sources that are consistent with the locus of AGN on the rest-frame $g-z$ color vs. [NeIII]$λ$3869 / [OII]$λλ$3726+3729 line ratio diagnostic diagram. We also use empirical AGN and galaxy templates to model the spectral energy distributions (SEDs) for the objects in our sample, and find that while there is significant variation in the observed SEDs for these objects, the majority require a strong AGN component. Finally, we use the results from our analysis of the optical spectra and the SEDs to compare our selection criteria to alternate criteria presented in the literature. These results verify the efficacy of selecting luminous obscured AGNs based on their WISE colors.

astro-ph.GA

Gravitational-wave Science in the High School Classroom

This article describes a set of curriculum modifications designed to integrate gravitational-wave science into a high school physics or astronomy curriculum. Gravitational-wave scientists are on the verge of being able to detect extreme cosmic events, like the merger of two black holes, happening hundreds of millions of light years away. Their work has the potential to propel astronomy into a new era by providing an entirely new means of observing astronomical phenomena. Gravitational-wave science encompasses astrophysics, physics, engineering, and quantum optics. As a result, this curriculum exposes students to the interdisciplinary nature of science. It also provides an authentic context for students to learn about astrophysical sources, data analysis techniques, cutting-edge detector technology, and error analysis.

physics.ed-ph