SearcharxivSearch

arXiv subjects

Lin Yan

Publications and source records attributed to Lin Yan.

At least 55 records · Page 3Linked to original sources

Halfway to the Peak: ice absorption bands at $z\approx0.5$ with JWST MIRI/MRS

This paper presents the first combined detections of CO$_2$, CO, XCN and water ices beyond the local Universe. We find gas-phase CO in addition to the solid phase CO. Our source, SSTXFLS J172458.3+591545, is a $z=0.494$ star-forming galaxy which also hosts a deeply obscured AGN. The profiles of its ice features are consistent with those of other Galactic and local galaxy sources and the implied ice mantle composition is similar to that of even more obscured sources. The ice features indicate the presence of a compact nucleus in our galaxy and allow us to place constraints on its density and temperature ($n>10^5$cm$^{-3}$ and $T=20-90K$). We infer the visual extinction towards this nucleus to be $A_V\approx6-7$. An observed plot of $τ_{Si}$ vs. $τ_{CO2}/τ_{Si}$ can be viewed as a probe for both the total dustiness of a system as well as the clumpiness of the dust along the line of sight. This paper highlights the potential of using {\sl JWST} MIRI spectra to study the dust composition and geometric distribution of sources beyond the local Universe.

astro-ph.GA

What's Behind PPO's Collapse in Long-CoT? Value Optimization Holds the Secret

Reinforcement learning (RL) is pivotal for enabling large language models (LLMs) to generate long chains of thought (CoT) for complex tasks like math and reasoning. However, Proximal Policy Optimization (PPO), effective in many RL scenarios, fails in long CoT tasks. This paper identifies that value initialization bias and reward signal decay are the root causes of PPO's failure. We propose Value-Calibrated PPO (VC-PPO) to address these issues. In VC-PPO, the value model is pretrained to tackle initialization bias, and the Generalized Advantage Estimation (GAE) computation is decoupled between the actor and critic to mitigate reward signal decay. Experiments on the American Invitational Mathematics Examination (AIME) show that VC-PPO significantly boosts PPO performance. Ablation studies show that techniques in VC-PPO are essential in enhancing PPO for long CoT tasks.

cs.LG

A General Framework for Augmenting Lossy Compressors with Topological Guarantees

Topological descriptors such as contour trees are widely utilized in scientific data analysis and visualization, with applications from materials science to climate simulations. It is desirable to preserve topological descriptors when data compression is part of the scientific workflow for these applications. However, classic error-bounded lossy compressors for volumetric data do not guarantee the preservation of topological descriptors, despite imposing strict pointwise error bounds. In this work, we introduce a general framework for augmenting any lossy compressor to preserve the topology of the data during compression. Specifically, our framework quantifies the adjustments (to the decompressed data) needed to preserve the contour tree and then employs a custom variable-precision encoding scheme to store these adjustments. We demonstrate the utility of our framework in augmenting classic compressors (such as SZ3, TTHRESH, and ZFP) and deep learning-based compressors (such as Neurcomp) with topological guarantees.

cs.DC

Flaming-hot Initiation with Regular Execution Sampling for Large Language Models

Since the release of ChatGPT, large language models (LLMs) have demonstrated remarkable capabilities across various domains. A key challenge in developing these general capabilities is efficiently sourcing diverse, high-quality data. This becomes especially critical in reasoning-related tasks with sandbox checkers, such as math or code, where the goal is to generate correct solutions to specific problems with higher probability. In this work, we introduce Flaming-hot Initiation with Regular Execution (FIRE) sampling, a simple yet highly effective method to efficiently find good responses. Our empirical findings show that FIRE sampling enhances inference-time generation quality and also benefits training in the alignment stage. Furthermore, we explore how FIRE sampling improves performance by promoting diversity and analyze the impact of employing FIRE at different positions within a response.

cs.LG

Enhancing Multi-Step Reasoning Abilities of Language Models through Direct Q-Function Optimization

Reinforcement Learning (RL) plays a crucial role in aligning large language models (LLMs) with human preferences and improving their ability to perform complex tasks. However, current approaches either require significant computational resources due to the use of multiple models and extensive online sampling for training (e.g., PPO) or are framed as bandit problems (e.g., DPO, DRO), which often struggle with multi-step reasoning tasks, such as math problem solving and complex reasoning that involve long chains of thought. To overcome these limitations, we introduce Direct Q-function Optimization (DQO), which formulates the response generation process as a Markov Decision Process (MDP) and utilizes the soft actor-critic (SAC) framework to optimize a Q-function directly parameterized by the language model. The MDP formulation of DQO offers structural advantages over bandit-based methods, enabling more effective process supervision. Experimental results on two math problem-solving datasets, GSM8K and MATH, demonstrate that DQO outperforms previous methods, establishing it as a promising offline reinforcement learning approach for aligning language models.

cs.LG

Process Supervision-Guided Policy Optimization for Code Generation

Reinforcement learning (RL) with unit test feedback has enhanced large language models' (LLMs) code generation, but relies on sparse rewards provided only after complete code evaluation, limiting learning efficiency and incremental improvements. When generated code fails all unit tests, no learning signal is received, hindering progress on complex tasks. To address this, we propose a Process Reward Model (PRM) that delivers dense, line-level feedback on code correctness during generation, mimicking human code refinement and providing immediate guidance. We explore various strategies for training PRMs and integrating them into the RL framework, finding that using PRMs both as dense rewards and for value function initialization significantly boosts performance. Our experimental results also highlight the effectiveness of PRMs in enhancing RL-driven code generation, especially for long-horizon scenarios.

cs.AI

Discovery of a years-delayed radio flare from an unusually slow-evolved tidal disruption event

SDSS J1115+0544 is a unique low-ionization nuclear emission-line region (LINER) galaxy with energetic ultraviolet (UV), optical and mid-infrared outbursts occurring in its nucleus. We present the results from an analysis of multi-wavelength photometric and radio follow-up observations covering a period of ~9 years since its discovery. We find that following a luminosity plateau of ~500 days, the UV/optical emission has decayed back to the pre-outburst level, suggesting that the nuclear outburst might be caused by a stellar tidal disruption event (TDE). In this case, SDSS J1115+0544 could be an unusually slow-evolved optical TDE with longest rise and decline time-scales ever found. Three years later than the optical peak, a delayed radio brightening was found with a 5.5 GHz luminosity as high as ~1.9x10^39 erg/s. Using a standard equipartition analysis, we find the outflow powering the radio emission was launched at t~1260 days with a velocity of beta<~0.1 and kinetic energy of E_K~>10^50 erg. The delayed radio brightening coupled with the disappearing plateau in the UV/optical light curves is consistent with the scenario involving delayed ejection of an outflow from a state transition in the disk. SDSS J1115+0544 is the first TDE displaying both a short-lived UV/optical plateau emission and a late-time radio brightening. Future radio observations of these TDEs in the post-plateau decay phase will help to establish the connection between outflow launching and changes in accretion rate.

astro-ph.HE

ZTF SN Ia DR2: Overview

We present the first homogeneous release of several thousand Type Ia supernovae (SNe Ia), all having spectroscopic classification, and spectroscopic redshifts for half the sample. This release, named the "DR2", contains 3628 nearby (z < 0.3) SNe Ia discovered, followed and classified by the Zwicky Transient Facility survey between March 2018 and December 2020. Of these, 3000 have good-to-excellent sampling and 2667 pass standard cosmology light-curve quality cuts. This release is thus the largest SN Ia release to date, increasing by an order of magnitude the number of well characterized low-redshift objects. With the "DR2", we also provide a volume-limited (z < 0.06) sample of nearly a thousand SNe Ia. With such a large, homogeneous and well controlled dataset, we are studying key current questions on SN cosmology, such as the linearity SNe Ia standardization, the SN and host dependencies, the diversity of the SN Ia population, and the accuracy of the current light-curve modeling. These, and more, are studied in detail in a series of articles associated with this release. Alongside the SN Ia parameters, we publish our force-photometry gri-band light curves, 5138 spectra, local and global host properties, observing logs, and a python tool to ease use and access of these data. The photometric accuracy of the "DR2" is not yet suited for cosmological parameter inference, which will follow as "DR2.5" release. We nonetheless demonstrate that the multi-thousand SN Ia Hubble Diagram has a typical 0.15 mag scatter.

astro-ph.CO

Halfway to the Peak: The JWST MIRI 5.6 micron number counts and source population

We present an analysis of eight JWST Mid-Infrared Instrument (MIRI) 5.6 micron images with $5\,σ$ depths of ~0.1 uJy. We detect 2854 sources within our combined area of 18.4 square arcminutes. We compute the MIRI 5.6um number counts including an analysis of the field-to-field variation. Compared to earlier published MIRI 5.6 um counts, our counts have a more pronounced knee, at roughly 2 uJy. The location and amplitude of the counts at the knee are consistent with the Cowley et al. (2018) model predictions, although these models tend to overpredict the counts below the knee. In areas of overlap, 84% of the MIRI sources have a counterpart in the COSMOS2020 catalog. These MIRI sources have redshifts that are mostly in the $z\sim0.5-2$, with a tail out to $z\sim5$. They are predominantly moderate to low stellar masses ($10^8-10^{10}$M$_{\odot}$) main sequence star-forming galaxies, suggesting that with ~2hr exposures, MIRI can reach well below $M^*$ at cosmic noon and reach higher mass systems out to $z\sim5$. Nearly 70% of the COSMOS2020 sources in areas of overlap now have a data point at 5.6um (rest-frame near-IR at cosmic noon) which allows for more accurate stellar population parameter estimates. Finally, we discover 31 MIRI-bright sources not present in COSMOS2020. A cross-match with IRAC channel 1 suggests that 10-20% of these are likely lower mass (M$_*\approx10^9$M$_{\odot}$), $z\sim1$ dusty galaxies. The rest (80--90%) are consistent with more massive but still very dusty galaxies at $z>3$.

astro-ph.GA

Optical and Radio Analysis of Systematically Classified Broad-lined Type Ic Supernovae from the Zwicky Transient Facility

We study a magnitude-limited sample of 36 Broad-lined Type Ic Supernovae (SNe Ic-BL) from the Zwicky Transient Facility Bright Transient Survey (detected between March 2018 and August 2021), which is the largest systematic study of SNe Ic-BL done in literature thus far. We present the light curves (LCs) for each of the SNe, and analyze the shape of the LCs to derive empirical parameters, along with the explosion epochs for every event. The sample has an average absolute peak magnitude in the r band of $M_r^{max}$ = -18.51 $\pm$ 0.15 mag. Using spectra obtained around peak light, we compute expansion velocities from the Fe II 5169 Angstrom line for each event with high enough signal-to-noise ratio spectra, and find an average value of $v_{ph}$ = 16,100 $\pm$ 1,100 km $s^{-1}$. We also compute bolometric LCs, study the blackbody temperature and radii evolution over time, and derive the explosion properties of the SNe. The explosion properties of the sample have average values of $M_{Ni}$ = $0.37_{-0.06}^{+0.08}$ solar masses, $M_{ej}$ = $2.45_{-0.41}^{+0.47}$ solar masses, and $E_K$= $4.02_{-1.00}^{+1.37} \times 10^{51}$ erg. Thirteen events have radio observations from the Very Large Array, with 8 detections and 5 non-detections. We find that the populations that have radio detections and radio non-detections are indistinct from one another with respect to their optically-inferred explosion properties, and there are no statistically significant correlations present between the events' radio luminosities and optically-inferred explosion properties. This provides evidence that the explosion properties derived from optical data alone cannot give inferences about the radio properties of SNe Ic-BL, and likely their relativistic jet formation mechanisms.

astro-ph.HE

A cosmic formation site of silicon and sulphur revealed by a new type of supernova explosion

The cores of stars are the cosmic furnaces where light elements are fused into heavier nuclei. The fusion of hydrogen to helium initially powers all stars. The ashes of the fusion reactions are then predicted to serve as fuel in a series of stages, eventually transforming massive stars into a structure of concentric shells. These are composed of natal hydrogen on the outside, and consecutively heavier compositions inside, predicted to be dominated by helium, carbon/oxygen, oxygen/neon/magnesium, and oxygen/silicon/sulphur. Silicon and sulphur are fused into inert iron, leading to the collapse of the core and either a supernova explosion or the direct formation of a black hole. Stripped stars, where the outer hydrogen layer has been removed and the internal He-rich layer (in Wolf-Rayet WN stars) or even the C/O layer below it (in Wolf-Rayet WC/WO stars) are exposed, provide evidence for this shell structure, and the cosmic element production mechanism it reflects. The types of supernova explosions that arise from stripped stars embedded in shells of circumstellar material (most notably Type Ibn supernovae from stars with outer He layers, and Type Icn supernovae from stars with outer C/O layers) confirm this scenario. However, direct evidence for the most interior shells, which are responsible for the production of elements heavier than oxygen, is lacking. Here, we report the discovery of the first-of-its-kind supernova arising from a star peculiarly stripped all the way to the silicon and sulphur-rich internal layer. Whereas the concentric shell structure of massive stars is not under debate, it is the first time that such a thick, massive silicon and sulphur-rich shell, expelled by the progenitor shortly before the SN explosion, has been directly revealed.

astro-ph.HE

Probing pre-supernova mass loss in double-peaked Type Ibc supernovae from the Zwicky Transient Facility

Eruptive mass loss of massive stars prior to supernova (SN) explosion is key to understanding their evolution and end fate. An observational signature of pre-SN mass loss is the detection of an early, short-lived peak prior to the radioactive-powered peak in the lightcurve of the SN. This is usually attributed to the SN shock passing through an extended envelope or circumstellar medium (CSM). Such an early peak is common for double-peaked Type IIb SNe with an extended Hydrogen envelope but is uncommon for normal Type Ibc SNe with very compact progenitors. In this paper, we systematically study a sample of 14 double-peaked Type Ibc SNe out of 475 Type Ibc SNe detected by the Zwicky Transient Facility. The rate of these events is ~ 3-9 % of Type Ibc SNe. A strong correlation is seen between the peak brightness of the first and the second peak. We perform a holistic analysis of this sample's photometric and spectroscopic properties. We find that six SNe have ejecta mass less than 1.5 Msun. Based on the nebular spectra and lightcurve properties, we estimate that the progenitor masses for these are less than ~ 12 Msun. The rest have an ejecta mass > 2.4 Msun and a higher progenitor mass. This sample suggests that the SNe with low progenitor masses undergo late-time binary mass transfer. Meanwhile, the SNe with higher progenitor masses are consistent with wave-driven mass loss or pulsation-pair instability-driven mass loss simulations.

astro-ph.HE

SN 2023zaw: an ultra-stripped, nickel-poor supernova from a low-mass progenitor

We present SN 2023zaw $-$ a sub-luminous ($\mathrm{M_r} = -16.7$ mag) and rapidly-evolving supernova ($\mathrm{t_{1/2,r}} = 4.9$ days), with the lowest nickel mass ($\approx0.002$ $\mathrm{M_\odot}$) measured among all stripped-envelope supernovae discovered to date. The photospheric spectra are dominated by broad He I and Ca NIR emission lines with velocities of $\sim10\ 000 - 12\ 000$ $\mathrm{km\ s^{-1}}$. The late-time spectra show prominent narrow He I emission lines at $\sim$1000$\ \mathrm{km\ s^{-1}}$, indicative of interaction with He-rich circumstellar material. SN 2023zaw is located in the spiral arm of a star-forming galaxy. We perform radiation-hydrodynamical and analytical modeling of the lightcurve by fitting with a combination of shock-cooling emission and nickel decay. The progenitor has a best-fit envelope mass of $\approx0.2$ $\mathrm{M_\odot}$ and an envelope radius of $\approx50$ $\mathrm{R_\odot}$. The extremely low nickel mass and low ejecta mass ($\approx0.5$ $\mathrm{M_\odot}$) suggest an ultra-stripped SN, which originates from a mass-losing low mass He-star (ZAMS mass $<$ 10 $\mathrm{M_\odot}$) in a close binary system. This is a channel to form double neutron star systems, whose merger is detectable with LIGO. SN 2023zaw underscores the existence of a previously undiscovered population of extremely low nickel mass ($< 0.005$ $\mathrm{M_\odot}$) stripped-envelope supernovae, which can be explored with deep and high-cadence transient surveys.

astro-ph.HE

MSz: An Efficient Parallel Algorithm for Correcting Morse-Smale Segmentations in Error-Bounded Lossy Compressors

This research explores a novel paradigm for preserving topological segmentations in existing error-bounded lossy compressors. Today's lossy compressors rarely consider preserving topologies such as Morse-Smale complexes, and the discrepancies in topology between original and decompressed datasets could potentially result in erroneous interpretations or even incorrect scientific conclusions. In this paper, we focus on preserving Morse-Smale segmentations in 2D/3D piecewise linear scalar fields, targeting the precise reconstruction of minimum/maximum labels induced by the integral line of each vertex. The key is to derive a series of edits during compression time; the edits are applied to the decompressed data, leading to an accurate reconstruction of segmentations while keeping the error within the prescribed error bound. To this end, we developed a workflow to fix extrema and integral lines alternatively until convergence within finite iterations; we accelerate each workflow component with shared-memory/GPU parallelism to make the performance practical for coupling with compressors. We demonstrate use cases with fluid dynamics, ocean, and cosmology application datasets with a significant acceleration with an NVIDIA A100 GPU.

cs.DC

WTP19aalnxx: Discovery of a bright mid-infrared transient in the emerging class of low luminosity supernovae revealed by delayed circumstellar interaction

While core-collapse supernovae (SNe) often show early and consistent signs of circumstellar (CSM) interaction, some exhibit delayed signatures due to interaction with distant material around the progenitor star. Here we present the discovery in NEOWISE data of WTP19aalnxx, a luminous mid-infrared (IR) transient in the outskirts of the galaxy KUG 0022-007 at $\approx 190$ Mpc. First detected in 2018, WTP19aalnxx reaches a peak absolute (Vega) magnitude of $\approx-22$ at $4.6 \, μ$m in $\approx3$ yr, comparable to the most luminous interacting SNe. Archival data reveal a $\gtrsim 5\times$ fainter optical counterpart detected since 2015, while follow-up near-IR observations in 2022 reveal an extremely red ($Ks-W2 \approx 3.7$ mag) active transient. Deep optical spectroscopy confirm strong CSM interaction signatures via intermediate-width Balmer emission lines and coronal metal lines. Modeling the broadband spectral energy distribution, we estimate the presence of $\gtrsim 10^{-2}$ M$_\odot$ of warm dust, likely formed in the shock interaction region. Together with the lack of nebular Fe emission, we suggest that WTP19aalnxx is a missed, low (optical) luminosity SN in an emerging family of core-collapse SNe distinguished by their CSM-interaction-powered mid-IR emission that outshines the optical bands. Investigating the Zwicky Transient Facility sample of SNe in NEOWISE data, we find $17$ core-collapse SNe ($\gtrsim 3$% in a volume-limited sample) without early signs of CSM interaction that exhibit delayed IR brightening, suggestive of dense CSM shells at $\lesssim 10^{17}$cm. We suggest that synoptic IR surveys offer a new route to revealing late-time CSM interaction and the prevalence of intense terminal mass loss in massive stars.

astro-ph.HE

Neutrino follow-up with the Zwicky Transient Facility: Results from the first 24 campaigns

The Zwicky Transient Facility (ZTF) performs a systematic neutrino follow-up program, searching for optical counterparts to high-energy neutrinos with dedicated Target-of-Opportunity (ToO) observations. Since first light in March 2018, ZTF has taken prompt observations for 24 high-quality neutrino alerts from the IceCube Neutrino Observatory, with a median latency of 12.2 hours from initial neutrino detection. From two of these campaigns, we have already reported tidal disruption event (TDE) AT 2019dsg and likely TDE AT 2019fdr as probable counterparts, suggesting that TDEs contribute >7.8% of the astrophysical neutrino flux. We here present the full results of our program through to December 2021. No additional candidate neutrino sources were identified by our program, allowing us to place the first constraints on the underlying optical luminosity function of astrophysical neutrino sources. Transients with optical absolutes magnitudes brighter that $-21$ can contribute no more than 87% of the total, while transients brighter than $-22$ can contribute no more than 58% of the total, neglecting the effect of extinction and assuming they follow the star formation rate. These are the first observational constraints on the neutrino emission of bright populations such as superluminous supernovae. None of the neutrinos were coincident with bright optical AGN flares comparable to that observed for TXS 0506+056/IC170922A, with such optical blazar flares producing no more than 26% of the total neutrino flux. We highlight the outlook for electromagnetic neutrino follow-up programs, including the expected potential for the Rubin Observatory.

astro-ph.HE

Establishing accretion flares from massive black holes as a source of high-energy neutrinos

The origin of cosmic high-energy neutrinos remains largely unexplained. For high-energy neutrino alerts from IceCube, a coincidence with time-variable emission has been seen for three different types of accreting black holes: (1) a gamma-ray flare from a blazar (TXS 0506+056), (2) an optical transient following a stellar tidal disruption event (TDE; AT2019dsg), and (3) an optical outburst from an active galactic nucleus (AGN; AT2019fdr). For the latter two sources, infrared follow-up observations revealed a powerful reverberation signal due to dust heated by the flare. This discovery motivates a systematic study of neutrino emission from all supermassive black hole with similar dust echoes. Because dust reprocessing is agnostic to the origin of the outburst, our work unifies TDEs and high-amplitude flares from AGN into a population that we dub accretion flares. Besides the two known events, we uncover a third flare that is coincident with a PeV-scale neutrino (AT2019aalc). Based solely on the optical and infrared properties, we estimate a significance of 3.6$σ$ for this association of high-energy neutrinos with three accretion flares. Our results imply that at least ~10% of the IceCube high-energy neutrino alerts could be due to accretion flares. This is surprising because the sum of the fluence of these flares is at least three orders of magnitude lower compared to the total fluence of normal AGN. It thus appears that the efficiency of high-energy neutrino production in accretion flares is increased compared to non-flaring AGN. We speculate that this can be explained by the high Eddington ratio of the flares.

astro-ph.HE

Long-rising Type II Supernovae in the Zwicky Transient Facility Census of the Local Universe

SN 1987A was an unusual hydrogen-rich core-collapse supernova originating from a blue supergiant star. Similar blue supergiant explosions remain a small family of events, and are broadly characterized by their long rises to peak. The Zwicky Transient Facility (ZTF) Census of the Local Universe (CLU) experiment aims to construct a spectroscopically complete sample of transients occurring in galaxies from the CLU galaxy catalog. We identify 13 long-rising (>40 days) Type II supernovae from the volume-limited CLU experiment during a 3.5 year period from June 2018 to December 2021, approximately doubling the previously known number of these events. We present photometric and spectroscopic data of these 13 events, finding peak r-band absolute magnitudes ranging from -15.6 to -17.5 mag and the tentative detection of Ba II lines in 9 events. Using our CLU sample of events, we derive a long-rising Type II supernova rate of $1.37^{+0.26}_{-0.30}\times10^{-6}$ Mpc$^{-3}$ yr$^{-1}$, $\approx$1.4% of the total core-collapse supernova rate. This is the first volumetric rate of these events estimated from a large, systematic, volume-limited experiment.

astro-ph.HE