SearcharxivSearch

arXiv subjects

Xiaohui Fan

Publications and source records attributed to Xiaohui Fan.

At least 19 recordsLinked to original sources

Learning JWST. I. A Foundation Model for New Population Discoveries and Morphology-Aware Photometric Redshift Measurements in the JADES Survey

We present FM-JADES-v1, a self-supervised foundation model for James Webb Space Telescope ({\em JWST}) deep-field science, trained with 482,444 objects from the {\em JWST} Advanced Deep Extragalactic Survey (JADES) Data Release 5 using multi-band imaging and the photometric catalog. The shared embedding space is trained without class labels. We demonstrate that FM-JADES-v1 can serve as a powerful tool for object discovery and improving property measurements using two experiments, blind active discovery and few-band photometric redshift. For blind object discovery, FM-JADES-v1 identifies rare object populations such as high-redshift galaxies and Little Red Dots (LRDs) without any prior population labels or population-specific selection criteria. These rare populations emerge as isolated islands in the embedding space, which can be identified without prior astrophysical knowledge. For few-band photometric redshift, FM-JADES-v1's learned embeddings achieve $\sigma_{\rm NMAD}=0.157$ in a strictly controlled three-band (F115W/F200W/F356W) photo-$z$ benchmark, compared to $\sigma_{\rm NMAD}= 0.44$ for template fitting. These results demonstrate the potential of self-supervised multi-modal representations as scalable discovery spaces for large astronomical surveys. Applied to ongoing and future wide-field surveys from JWST, Roman, Euclid, and Rubin/LSST, this framework could enable systematic searches for rare populations, as well as enabling multiple downstream tasks such as improving astrophysical property measurements.

astro-ph.GA

The Roman eXtreme Deep Field (RXDF)

The Roman eXtreme Deep Field (RXDF) program is one of the five General Astrophysics Survey (GAS) programs approved for observing time with the Nancy Grace Roman Space Telescope in Cycles 1 and 2. It has been allocated 386.41 hours to carry out an imaging survey to AB = 30 mag (5-sigma) over ~140x larger area than the Hubble eXtreme Deep Field (HXDF) full-depth area (ACS+WFC3/IR). The RXDF will cover the full Roman wavelength range with 7 bands, reaching AB = 30 mag in RZYJH, 29 mag in F, and 28 mag in K, over a full-depth area of 678.75 arcmin^2 embedded in a total area of 1,243 arcmin^2, and far exceeding the depths of the Roman Core Community Surveys (CCS). The RXDF is within the Euclid Ultra Deep Field (EUDF) near the North Ecliptic Pole (NEP), a strategic long-term field for generational space facilities, with a wealth of multi-wavelength data including extensive coverage from the James Webb Space Telescope (JWST) NEXUS Treasury program. The observations will cover 3 epochs at a 1-year cadence, each epoch divided into 3 sub-epochs ~10 days apart, enabling time-domain studies on time baselines from ~10 days to over ~2 years. The RXDF is uniquely positioned to address critical questions in reionization, large scale structure (LSS), growth of supermassive black holes (SMBHs), little red dots (LRDs), and high-z supernovae (SNe); the volumes probed by HST+JWST are too small at these extreme depths, and even the deepest CCS tiers are too shallow. In addition to our key objectives, a wealth of additional science will be enabled by engaging the community with our rapidly released datasets, revolutionizing a wide range of science for a lasting legacy. This short document, which is converted from the approved RXDF proposal, aims to provide the community with a summary of the program.

astro-ph.GA

NEXUS: Transient Searches and First Results from Year One Observations

We describe ongoing efforts of high-redshift transient searches using multi-epoch NIRCam imaging (F200W+F444W) and NIRSpec MSA/PRISM spectroscopy from the NEXUS JWST multi-cycle Treasury program targeting the north ecliptic pole region. The transient search area covers $\sim 61.5\,{\rm arcmin^2}$ between the reference epoch and each subsequent NEXUS-Deep epoch at a cadence of $\sim2$~months. In the first year of observations from NEXUS, we detect 68 robust transients, with the host photometric redshift distribution declining rapidly at $z>2$ but extending to $z_{\rm phot}\approx 6$. In addition, we obtained secure spectroscopic redshifts for 37 transients ($\sim54\%$) from NIRSpec/PRISM and NIRCam/WFSS, with seventeen at $1 < z < 2$, eight at $2 < z < 3$, two at $3 < z < 4$, and one tentative host association at z = 6.151. Overall, the NEXUS program recovers observed supernova (SN) rates broadly consistent with other SN search programs with JWST. While NEXUS achieves the highest transient detection efficiency, 6.6 SNe per imaging hr ($2.6\times$ COSMOS-SN and $26\times$ JADES-SN), the limited filter coverage (F200W+F444W only) limits robust identification and classification of high-$z$ SNe for deep spectroscopic follow-up. We describe the details of the data reduction and transient detection pipeline, and report the Year-1 transient sample along with their light curves, available MSA spectroscopy, host association, and the raw detection rate. We also describe and release a new PSF photometry package that properly accounts for correlated pixel noise from combining drizzled images.

astro-ph.HE

Unraveling the Real Working Mechanism and Inherent Flaws of GAE: A Method for Interpreting Transformer Processes from an Economic Perspective

We observe a phenomenon that current algorithmic research in the field of explainable artificial intelligence primarily pursues better performance on several proxy metrics. On the one hand, these proxy metrics themselves are more or less flawed and cannot properly measure the quality of methods. On the other hand, metric-oriented research approaches often lead to the neglect of the rationality and interpretability of the methods themselves. Explainable artificial intelligence is abbreviated as XAI. The metric-driven research paradigm has resulted in a lack of interpretability of the relevant XAI methods themselves. Accordingly, there is a need for interpretability research on XAI methods, which can be playfully referred to as XXAI. This paper is one of our works on XXAI. This paper takes Generic Attention-model Explainability (GAE), a widely influential model interpretation method , or rather, XAI method that represents an important technical route, as the research object, and explores the real working mechanism and flaws of this method as well as the technical route it represents. Based on the conclusions of this study, it may be necessary to re-examine or verify GAE-related methods and their domain applications. We argue that GAE is an interpretation method that focuses on the attention process. After pointing out the working mechanism and flaws of GAE, we propose Cumulative Asset Holdings (CAH), a more reasonable Transformer interpretation method integrating both process-based and feature-based ideas from an economic zero-sum games perspective. In addition, it is worth noting that our method is applicable to models with special tokens, where existing methods may suffer from limitations. The model simplification research method and the analysis of additive operations adopted in this study may provide inspiration for other research works in XAI.

cs.AI

Low Ly$\alpha$ Visibility in Galaxy Overdensities: Reionization Topology and Neutral-Fraction Ceilings from DIVER over $4.8<z<11$

Ly-alpha emission is widely used to trace cosmic reionization, but its interpretation depends on how Ly-alpha visibility varies with galaxy environment. We use deep JWST/NIRSpec observations from Deep Insights into UV Spectroscopy at the Epoch of Reionization (DIVER) in GOODS-N to measure Ly-alpha visibility for 250 galaxies at 4.8 25 A. We combine these measurements with H-alpha and [O III] emitters from JWST/NIRCam wide-field slitless spectroscopy to map the density field around each DIVER galaxy. Galaxies with high Ly-alpha equivalent widths (W_Lyalpha>25 A) or high effective Ly-alpha escape fractions (f_esc,Lyalpha^eff>0.05) tend to lie farther from nearby H-alpha and [O III] emitters than galaxies with lower Ly-alpha visibility. The clearest signal occurs near the prominent GOODS-N overdensity at z~5.2, where fewer than 15% of galaxies show strong Ly-alpha emission. This trend is opposite to the simplest inside-out reionization expectation that overdensities produce larger ionized regions and enhance Ly-alpha visibility. Possible explanations include circumgalactic and local intergalactic opacity, dense absorbers, and gas kinematics. We also derive an empirical upper envelope for f_esc,Lyalpha^eff and calibrate it with reionization simulations. Interpreting this envelope as a limiting IGM-attenuation signal gives neutral-fraction ceilings of _max=0.36, 0.76, 0.74, 0.84, and 1.0 at z~5.2, 5.8, 6.7, 7.7, and 9.8, respectively. The z~8 ceiling disfavors an almost completely neutral IGM at this epoch. These results support patchy reionization already underway by z~8 and show that galaxy Ly-alpha visibility encodes both large-scale ionization topology and near-source gas structure.

astro-ph.GA

Quasar Impostors: Two Extremely UV-Bright ($M_{\rm UV}\approx-23.5$) Reionisation-Epoch Galaxies Powered by Very Massive Stars

The extreme bright end of the galaxy UV luminosity function during reionisation remains poorly constrained, particularly where the galaxy and quasar luminosity functions overlap and source classification becomes ambiguous. We present JWST/NIRSpec and ALMA Band-6 observations of J1450-0144 ($z=6.627$) and J1429-0104 ($z=6.796$), two $M_{\rm UV}\simeq-23.5$ sources originally classified as faint quasars by SHELLQs. NIRSpec reveals blue UV continua, strong P Cygni profiles in N V, Si IV, and C IV, broad He II $\lambda1640$ emission with rest-frame equivalent widths of $8.8\pm1.2$ and $3.7\pm1.1$ {\AA}, respectively, and narrow nebular lines, reclassifying both as extremely UV-luminous galaxies. Standard population-synthesis models cannot simultaneously reproduce the strong He II and wind features, whereas models incorporating very massive stars (VMS; $M\gtrsim100\,M_\odot$) with dedicated wind prescriptions can. These models favor a star-formation duration of 2-4 Myr for J1450-0144, with a broader allowed range for J1429-0104, stellar masses of $\log(M_\star/M_\odot)\approx9.2$-$9.9$, and star-formation rates of $\simeq300$-$540\,M_\odot\,{\rm yr}^{-1}$. Under the same VMS wind models, equivalent-width diagnostics imply $M_{\rm up}\gtrsim225\,M_\odot$ for J1429-0104, while J1450-0144 lies beyond even the $M_{\rm up}=475\,M_\odot$ grid. ALMA detects luminous [C II] 158 $\mu$m emission in both systems, with $L_{\rm [CII]}\approx0.8$ and $4.1\times10^{9}\,L_\odot$, respectively. J1429-0104 additionally shows bright dust continuum, with both [C II] and dust offset by $\sim5.4$ kpc from its UV emission. These sources demonstrate that VMS can power some of the most UV-luminous galaxies at cosmic dawn and show that source classifications, and hence the inferred demographics of both galaxies and quasars in the crossover regime, need revisiting.

astro-ph.GA

Misaligned or chaotic? A strong break of axial symmetry in the local LRD J1025 revealed with VLT/FORS2 spectropolarimetry

Little Red Dots (LRDs) are compact active galactic nuclei (AGN) with unusual spectral energy distributions and broad Balmer emission, candidate signposts of rapid black-hole growth. We present VLT/FORS2 optical linear spectropolarimetry of the closest known LRD, SDSS J102530.29+140207.3, at z=0.1. In total light, we detect spatially extended narrow-H$\alpha$ emission, probably tracing the host galaxy. We measure a nearly grey continuum polarisation $p_{\rm cont}=1.53\pm0.04$(rand.)$\pm$0.20(syst.) per cent, while broad H$\alpha$ is less polarised, $p_{\rm H\alpha}=0.58-0.84$ per cent. We rule out polarisation by Milky Way dust and dilution by an unpolarised line. The polarised continuum with a depolarised line resembles local Seyfert-1 nuclei and favours a single dominant source over multi-source explanations. After Stokes-continuum subtraction, broad H$\alpha$ shows no blue-to-red swing, but there is a significant (48$\pm$4)$^\circ$ continuum-to-line offset in polarisation angle. This implies a break of axial symmetry inside this object, posing interesting geometrical challenges to all existing LRD models and frameworks. We discuss different possible origins of this symmetry break and possible paths to discriminate between them with future observations.

astro-ph.GA

The Twentieth Data Release of the Sloan Digital Sky Survey: First All-Sky BOSS Spectra, eROSITA-SDSS-V Mapper Coordinated Observations, and a Preview of the Local Volume Mapper

This paper presents the twentieth data release (DR20) from the Sloan Digital Sky Survey, the third data release of its fifth generation (SDSS-V). SDSS-V is a panoptic spectroscopy survey that is mapping the stars, gas, and galaxies through three scientific programs: the Milky Way Mapper (MWM), the Local Volume Mapper (LVM), and the Black Hole Mapper (BHM). DR20 presents the first optical (BOSS) SDSS-V spectra from southern hemisphere for the MWM and BHM surveys; new optical MWM and BHM data from the northern hemisphere are also available, for a total over 3 million spectra of 1.5 million stars and half a million galaxies and quasars, with galactic and extragalactic x-ray targets coordinate with eROSITA DR2. DR20 includes integral field spectroscopy maps from LVM of six targets and 169 tiles, spanning Galactic HII regions, planetary nebulae, and nearby galaxies. Additionally, eighteen value added catalogs are also released with DR20, based on SDSS-V MWM and BHM data, and we present a new LVM visualization tool including an RGB HiPS map as a value added product.

astro-ph.GA

A quasar hatching from a buried red phase at z = 3.7

We present JADES-GS 209777, previously cataloged as CANDELS J033238.02-274626.2, hereafter "the Hatchling," a red quasar at $z=3.711$. While the source has been reported in earlier deep-field catalogs, our multiwavelength analysis reveals a visible active nucleus still embedded in a dense gas- and dust-rich environment. Red quasar continua are often attributed to dust attenuation, including non-standard extinction curves, but the highly comprehensive multiwavelength data for this source provide direct constraints on the material being cleared. Using JWST/NIRSpec, NIRCam, MIRI, HST, MUSE, Chandra, ALMA, and VLA data, we detect broad emission lines and strong X-ray emission, showing that the active nucleus is at least partially exposed. We also detect H$\alpha$ and He I absorption, indicating dense gas close to the nucleus. Kinematically disturbed O I, Mg II, Na D, and [O III] features, together with extended Ly$\alpha$ emission over $\gtrsim 20$ kpc, further show that multiphase gas is being accelerated from the nuclear region into the host-galaxy environment. The ALMA detection reveals strong dust emission, with the inferred infrared luminosity placing the system in the ULIRG regime. The continuum is red and sharply declining toward the rest-frame UV, resembling compact red AGNs, and may reflect extreme dust attenuation, gas reprocessing, possible BAL-like absorption, or a combination of these effects. Regardless of which mechanism dominates the continuum shape, the line diagnostics show that the visible nucleus remains partially obscured by nearby material. The Hatchling therefore represents a unique opportunity to explore a poorly known transition phase in which feedback is likely clearing an enshrouded quasar and allowing it to emerge toward a more unobscured active nucleus.

astro-ph.GA

The MIRI Early Obscured-AGN Wide Survey (MEOW): A Population of Hidden AGN at $z \gtrsim 5$ Revealed by JWST/MIRI Imaging

We present the MIRI Early Obscured-AGN Wide Survey (MEOW), a JWST/MIRI imaging survey designed to detect dust-obscured active galactic nuclei (AGN) across cosmic time, with a particular focus on the high-redshift universe at $z \gtrsim 5$. MEOW observes the GOODS-N and GOODS-S fields with 43 pointings covering 95 arcmin$^2$ with the F1000W and F2100W filters, reaching depths of 0.5 and 3.6 $\mu$Jy ($5\sigma$), respectively. Using spectral energy distribution (SED) modeling combining MEOW photometry with archival HST, JWST/NIRCam, and SCUBA-2 data, we identify a sample of 16 MIRI-selected AGN at $z = 4.5$--$7.2$ (12 spectroscopically confirmed), spanning bolometric luminosities of $L_{\rm bol} = 10^{44.6}$--$10^{46.4}$~erg~s$^{-1}$. Twelve of the 16 AGN are newly identified in this work, including at least five narrow-line AGN representing the obscured population to which broad-line spectroscopic searches are insensitive. Two broad-line AGN exhibit markedly different mid-infrared emission properties, consistent with one being a little red dot (LRD) and the other either a typical AGN or an LRD with unusually strong hot-dust emission. The MIRI-selected AGN bolometric luminosity function at $z = 4.5$--$6$ yields number densities comparable to those of broad-line AGN and LRDs, suggesting that obscured AGN contribute significantly to the total AGN census at these epochs. The narrow-line AGN reside in diverse host environments, with evidence for both circumnuclear and host-galaxy-scale obscuration, pointing to multiple physical mechanisms at work. These results establish JWST/MIRI imaging as an indispensable component of a multi-faceted approach to a complete census of early supermassive black hole growth.

astro-ph.GA

First-star imprints in a metal-poor galaxy overdensity near the end of reionization

The first generation of stars, known as Population III (Pop III), formed from primordial gas consisting solely of hydrogen and helium and is believed to have emerged only a few hundred million years after the Big Bang. Detecting the chemical enrichment of metal-poor circumgalactic gas offers a promising way to trace the enrichment signature of Pop III stars. Along the sightline to the quasar SDSS J0100+2802, a metal absorber at $z = 5.945$, showing over-abundant carbon and silicon compared to solar, has been reported to be consistent with the enrichment pattern of Pop III stars. With the James Webb Space Telescope, we report the discovery of an unusually metal-poor galaxy overdensity of 17 members (mean metallicity $\approx 3\%$ solar) near this metal absorber, which is $\sim 0.4$ dex more metal-poor than coeval galaxies in similarly overdense environments. This less chemically evolved system may have provided favorable conditions for preserving the absorption signatures of Pop III enrichment. We infer a minimum dark matter halo of $\log(M_{\mathrm{h,min}}/M_{\odot})=10.68^{+0.93}_{-1.72}$, supporting late-time Pop III formation at the outskirts of atomic hydrogen cooling halos. Our findings open a promising observational pathway to identify the chemical imprints of the first stars and constrain the conditions for their formation.

astro-ph.GA

Probing Direct Contributions of Galaxies and AGN to Cosmic Reionization in a Quasar Field J0226+0302 with JWST NIRCam and NIRSpec

We present JWST Cycle 2 NIRCam and NIRSpec observations in a quasar field J0226+0302 at z=6.5412 to probe the direct connections between the intergalactic medium (IGM), galaxies, and AGN during reionization. This field was previously observed by the JWST ASPIRE program and eight [OIII]-emitting galaxies were detected at 5.3<z<6.4 with a single NIRCam pointing. Using new NIRCam and NIRSpec observations, we identify 65 additional line-emitting galaxies at 5.3<z<6.4. The IGM-galaxy cross-correlation function shows a ~2 sigma excess IGM transmission at ~10-40 cMpc from galaxies when compared with the average IGM transmission, suggesting a significant contribution from regions traced by star-forming galaxies to the local ionizing background during reionization. The IGM-galaxy cross-correlation function is consistent with THESAN simulations with an IGM neutral fraction of 5%-7% and an average ionizing photon escape fraction f_esc of 6% from galaxies. Among 49 line-emitting galaxies observed by NIRSpec, we identify four AGN through detection of broad H-alpha emission lines with an AGN fraction of (8+/-4)%. By measuring the IGM effective optical depth around the AGN and the IGM-AGN cross-correlation function, we find that the IGM transmission is higher within 5 cMpc/h of the AGN than around the majority of [OIII] emitters. We interpret the excess IGM transmission as resulting from the local radiation enhancement by the AGN, and estimate f_esc of 50%-100% of the AGN from the IGM-AGN cross-correlation function. Future JWST NIRSpec observations in quasar fields will yield a more constraining IGM-AGN cross-correlation function, providing further insights into the roles of galaxies and AGN in reionization.

astro-ph.GA

High-Resolution ALMA Imaging for a Gravitationally-lensed Quasar at $z=6.5$: Constraining the AGN Contribution to Galactic-Scale Dust Heating

We present high-resolution (beam size $0\farcs076\times0\farcs040$) Atacama Large Millimeter/submillimeter Array (ALMA) observations of the far-infrared $(\lambda_\text{rest}=162.7\mu\rm{m})$ dust continuum of J0439+1634, a gravitationally lensed quasar at $z=6.52$. We perform pixelated lens modeling for the visibility data, finding that J0439+1634 is well-described by a singular isothermal ellipsoid plus an external shear lensing model. The best-fit lensing potential exhibits a naked-cusp configuration, confirming the finding in Fan et al. (2019). The reconstructed source plane continuum emission shows a compact bright core, with size $\lesssim200$ pc and peak brightness $\sim0.6 \text{ Jy arcsec}^{-2}$. The total continuum flux at 245 GHz is $3.36\pm0.02$ mJy. The flux magnification is {$4.63\pm0.03$}, indicating an average source-plane resolution of $0\farcs019$ (equivalent to 104 pc). The spatial resolution around the supermassive black hole reaches $\sim36$ pc. %Using the new lensing model, we re-fit the Hubble Space Telescope image for J0439+1634, and find that the position of the optical quasar is consistent with the brightest pixel in the dust continuum map. Leveraging the exceptional source-plane resolution, we build a radiative transfer model to describe the observed dust emission profile. The best-fit model indicates that heated dust from the active galactic nucleus (AGN) dominates the sub-millimeter emission at $r\lesssim100$ pc and that star-heated dust dominates the outer region of the host galaxy. AGN heating contributes {$\sim13\%$} to the observed sub-mm flux. Therefore, previous far-infrared-based star formation rate measurements for most high-redshift quasars are likely mildly overestimated.

astro-ph.GA

A $z \sim$ 6.2 Quasar on the Local M$_{\rm BH}$-$\sigma_{\rm \ast}$ Relation Quenching Its Host Galaxy from the Aether Survey

We report JWST/NIRSpec integral field unit (IFU) observations of the quasar J1512$+$4422 at $z \sim 6.2$ from the Aether survey. At $\sim$900 Myr after the Big Bang, this object already lies on the $M_{\rm BH}$-$\sigma_\ast$ relation found in the local universe, with an $M_{\rm BH} \simeq 8.9\times10^8\,M_\odot$ and a stellar velocity dispersion $\sigma_\ast \simeq 288$ km s$^{-1}$. We detect an outflow with a velocity of $\sim$478 km s$^{-1}$ in the nuclear region, which likely extends to $\sim$3.2 kpc in projection and has a median velocity of $\sim$352 km s$^{-1}$. The outflow dynamical time scale ($\sim$ 9 Myr) is consistent with the time scale of the current quenching process based on the star formation history as reported previously. The total mass outflow rate (92.6$^{+92.6}_{-74.1}$ M$_{\odot}$ yr$^{-1}$) is larger than the current star formation rate (0.9$^{+3.8}_{-0.8}$ or 4.3$^{+5.8}_{-3.7}$ M$_{\odot}$ yr$^{-1}$), and the total kinetic energy outflow rate (0.6$^{+0.6}_{-0.5}$\% of quasar luminosity) meets the threshold for negative quasar feedback as suggested by simulations. These results suggest that the outflow is capable of suppressing/quenching the star formation activity within the host galaxy. Furthermore, J1512$+$4422 exhibits $\sigma_\ast$, stellar mass and size similar to those of $z \gtrsim$ 3 quiescent/post-starburst galaxies, implying a link between the two. Overall, for objects like J1512$+$4422, the evolution of their SMBHs and host galaxies appears to be tightly coupled within the first billion years. The quasar feedback likely plays a critical role in both placing them on the $M_{\rm BH}$--$\sigma_\ast$ relation and quenching.

astro-ph.GA

A Steep-Extinction Quasi-stellar Object at z=4.6: JWST Evidence for Abundant Small Dust Grains

The rapid accumulation of massive dust reservoirs in the early Universe remains a major challenge in astrophysics. While core-collapse supernovae can inject large dust grains ($a \gtrsim 0.1\,\mu{\rm m}$) on short timescales, explaining the total dust budgets in the early Universe likely requires efficient grain growth in the interstellar medium (ISM). Such growth depends critically on an abundant population of small grains, which maximize the surface area available for accretion and may be generated by rapid dust-processing or dust-formation channels. Here, we report the discovery of a QSO, UDS-27023, at $z=4.556\pm0.003$, identified using JWST/NIRSpec spectroscopy. By quantitatively comparing the spectra to QSO composite templates, we find that UDS-27023 displays an exceptionally steep far-UV extinction curve ($A_{1500}/A_V \approx 8$) but notably lacks the 2175 A bump ($A_\mathrm{bump}/A_V<0.34$ at $3\sigma$), indicating a dominance of small silicate dust grains. We interpret this phenomenology as evidence for active small-grain production and processing in the QSO environment. Mechanical shattering of pre-existing large grains by QSO-driven shocks and outflows provides one natural pathway, while in situ condensation of silicate grains inside dense QSO-driven winds may offer an additional route. Such a population of steep-extinction QSOs (SEQs) may therefore reveal a short-lived phase in which luminous active galactic nuclei generate, process, and redistribute small grains, potentially facilitating rapid ISM grain growth and enriching the circumgalactic medium.

astro-ph.GA

MAMMOTH-Grism: Gas-phase Metallicity Gradients of Star-forming Galaxies in Protocluster Environments at Cosmic Noon

Environment plays a crucial role in shaping galaxy formation, yet the impact of overdensities on the internal chemical structure of galaxies at cosmic noon is still under debate. Here, we present spatially resolved gas-phase metallicity gradients for 42 star-forming galaxies in three massive protoclusters at $z \sim 2.3$, derived fromHubble Space Telescope (HST) slitless grism spectroscopy from the MAMMOTH-Grism survey. We find that the majority (29 of 42, $\sim$69%) of these protocluster members exhibit positive (inverted) metallicity gradients, a fraction significantly higher than observed in field galaxies of similar mass and redshift. By examining correlations with global properties, we show that these positive gradients are strongly associated with galaxies that are metal-deficient relative to the field mass-metallicity relation, particularly among the massive population ($\log(M_*/M_\odot) > 9.95$). These trends suggest that galaxies in dense protocluster environments experience substantial, enhanced inflows of pristine gas toward their central regions, which dilute the central metallicity and produce the observed inverted gradients. Our results provide observational evidence that environmental effects actively regulate gas accretion and chemical redistribution during the peak epoch of cosmic star formation.

astro-ph.GA

Generic Interpretation Approach for Transformer Models Incorporating Heterogenous Attention Structures

Transformer has significantly propelled the development of artificial intelligence, and certainly the development of agents as well. We categorize attention structures of Transformer into two types based on the source of the input information: homogenous and heterogenous attention structures. Heterogenous attention structures, with co-attention as a typical example, process information from different sources. Heterogenous attention structure is the foundation for Transformer models to achieve more complex functions and integrate more modal information. Whether for research purposes or policy requirements, the interpretation of Transformer models with heterogenous attention structures is an important task. The fusion of information from different sources brings new challenges. Our work mainly includes two parts: method and experimentation. In terms of method, we propose an interpretation method for Transformer models with heterogenous attention structures. In terms of experimentation, based on our experimental analysis paradigm, we interpret the operating mechanisms of representative models, conduct semantic interpretation and logical interpretation.

cs.CV

The Neglected Baseline in Model Interpretation

We observe that existing model interpretation methods generally ignore the baseline, and such neglect often results in imprecise or even incorrect interpretation. In this paper, we reformulate the task of model interpretation and the interpretation principles for model interpretation results to demonstrate the importance of the baseline. We further unify gradient-based methods, Integrated Gradients (IG) methods, and Taylor expansion, clarifying the connections among them and explicitly identifying the baseline for each method. On this basis, we analyze the flaws and errors in related model interpretation methods (IG, LayerCAM, ODAM, Difference Map). We advocate evaluating the quality of model interpretation results precisely through the attribution error between the attribution result and the attribution target, rather than adopting flawed evaluation methods, such as those based on marginal-effect or the assumption of perfect model performance. We revise IG and develope a model interpretation method with a clear and reasonable baseline, achieving better results. Our method supports model interpretation based on features from any layer. Interpretation based on features from different layers are all reasonable, and the differences among these results reflect varying degrees of feature extraction at different feature extraction stages.

cs.CV