SearcharxivSearch

arXiv subjects

Xiaofan Ma

Publications and source records attributed to Xiaofan Ma.

3 recordsLinked to original sources

Deep learning with hybrid frequency differencing and principal component analysis for 21-cm foreground and beam mitigation

Twenty-one-centimeter intensity mapping is a powerful probe of the large-scale distribution of neutral hydrogen (HI) and cosmological observables such as baryon acoustic oscillations. A major challenge is contamination from bright foregrounds and frequency-dependent beam effects, which can lead to signal loss in traditional methods such as principal component analysis (PCA). We develop a hybrid approach that trains a U-shaped convolutional neural network (UNet) on two input channels derived from frequency differencing (FD) and PCA cleaning, enabling it to exploit their complementary behavior across different scales. This two-channel strategy achieves improved performance, maintaining the cross-correlation power spectrum close to unity on large scales under a cosine beam and improving by 5\%-8\% relative to either FD- or PCA-based UNet alone. We further show that the method can robustly recover the HI signal even when the beam model is imperfect and differs between training and testing, with the large-scale cross-correlation remaining close to unity within the $1\sigma$ level. These results demonstrate that the proposed approach provides a robust framework for HI signal reconstruction under realistic observational conditions.

astro-ph.CO

VisAug: Facilitating Speech-Rich Web Video Navigation and Engagement with Auto-Generated Visual Augmentations

The widespread adoption of digital technology has ushered in a new era of digital transformation across all aspects of our lives. Online learning, social, and work activities, such as distance education, videoconferencing, interviews, and talks, have led to a dramatic increase in speech-rich video content. In contrast to other video types, such as surveillance footage, which typically contain abundant visual cues, speech-rich videos convey most of their meaningful information through the audio channel. This poses challenges for improving content consumption using existing visual-based video summarization, navigation, and exploration systems. In this paper, we present VisAug, a novel interactive system designed to enhance speech-rich video navigation and engagement by automatically generating informative and expressive visual augmentations based on the speech content of videos. Our findings suggest that this system has the potential to significantly enhance the consumption and engagement of information in an increasingly video-driven digital landscape.

cs.MM

More Than Beautiful: Exploring Design Features, Practical Perspectives, and Implications of Artistic Data Visualization

Standing at the intersection of science and art, artistic data visualization has gained popularity in recent years and emerged as a significant domain. Despite more than a decade since the field's conceptualization, a noticeable gap remains in research concerning the design features of artistic data visualizations, the aesthetic goals they pursue, and their potential to inspire our community. To address these gaps, we analyzed 220 data artworks to understand their design paradigms and intents, and construct a design taxonomy to characterize their design techniques (e.g., sensation, interaction, narrative, physicality). We also conducted in-depth interviews with twelve data artists to explore their practical perspectives, such as their understanding of artistic data visualization and the challenges they encounter. In brief, we found that artistic data visualization is deeply rooted in art discourse, with its own distinctive characteristics in both inner pursuits and outer presentations. Based on our research, we outline seven prospective paths for future work.

cs.HC