SearcharxivSearch

arXiv subjects

Chen Xu

Publications and source records attributed to Chen Xu.

At least 19 recordsLinked to original sources

ConversationalVoice: Full-Duplex Speech Data from Real Conversations through Source-Faithful Reconstruction and Conversation-Grounded Expansion

Full-duplex speech models require training data that preserves turn-taking, overlap, interruption, and backchannel behavior, yet these signals are entangled across speakers in noisy real-world recordings. We present Conversational Voice, a pipeline that converts real two-speaker excerpts into three complementary training-data artifacts. (1) Separation recovers speaker-specific tracks with stable speaker assignments, a canonical transcript, and naturally observed interaction timing. (2) Reconstruction generates speech in matched voices from a fixed source transcript, reconstructs the source turn order, pauses, and overlaps, and adds word-level alignment and delivery instructions. (3) Expansion generates new dialogue constrained by the source context, speakers, and observed interaction pattern. Automatic speaker-verification metrics remain strong across stages, with same-speaker similarity of 0.983-0.991 and positive discrimination margins of 0.199-0.209. Predicted speech quality (NISQA MOS) is 3.56 for separation, 4.41 for reconstruction, and 4.61 for expansion. A Gemini-based automatic evaluator assigns expansion mean scores of 4.94/5 for contextual coherence and 4.80/5 for dialogue naturalness. Expansion and reconstruction exhibit broadly similar interaction profiles; expansion's turn, overlap-event, backchannel, and interruption rates are 4.6%, 8.0%, 13.2%, and 16.0% lower, respectively. We evaluate data properties only; downstream gains in full-duplex model training remain for future work.

cs.CL

An Elementary Proof of the $\widetilde O(n^{1/3})$ Bound for Separating Words

For two distinct binary words of length $n$, the separating words problem asks for a small deterministic finite automaton that accepts exactly one of them. Chase proved a $\widetilde O(n^{1/3})$ upper bound using a complex-analytic estimate for sparse polynomials. We replace that estimate by a finite-difference argument and a second-order real recurrence cutoff. The resulting elementary proof gives an explicit bound of $O(n^{1/3}(\log n)^{7/3})$ states.

cs.FL

Memory Anchors for Continual Robot Learning

Robot policies deployed in the wild should have the capability to continually learn new tasks without forgetting existing behaviors. A common approach to combat such catastrophic forgetting is to train on new task data with a replay buffer of previously learned task data. Although this buffer is commonly sampled randomly from all prior experiences, we show that a small set of these experiences contributes greatly in anchoring past performance. We call these experiences Memory Anchors. We identify Memory Anchors in regions where representations of new-task observations collapse onto those of old-task observations even though the tasks require conflicting actions, like when a familiar object must be manipulated in a new way. Rehearsing old data in this region plays a key role in preventing destructive overwriting of past task knowledge, serving as this critical Memory Anchor role. Excluding only 10% Memory Anchors before sampling the buffer leads to more than a 4.5x increase in catastrophic forgetting on the LIBERO benchmark suites. Conversely, enriching the replay buffer with Memory Anchors can decrease high-conflict task forgetting by 63% and enables successful continual learning of two task sequences on a real robot. Videos and additional visualizations can be found at https://robot-adaptation.github.io/MemoryAnchors

cs.RO

Mechanism Design for Generative Engines: From Exploitation toward Win-Win Outcomes

Generative engines are reshaping the web ecosystem by making citations a key mechanism for allocating attention, attribution, and downstream value. This creates a strategic tension: content providers are incentivized to optimize for model citation, while platforms must preserve answer quality and trustworthy attribution. We show that this tension can escalate into citation wars. In repeated simulations, state-of-the-art generative engine optimization (GEO) attacks adapt to conventional defenses by producing citation-seeking rewrites that degrade document quality and introduce unsupported claims. To study this problem, we formulate the supplier--platform interaction as a repeated Stackelberg game with partial monitoring. A local best-response analysis identifies when citation competition approaches an inert stationary outcome. Motivated by this finding, we propose a platform--creator mechanism called VCR based on verifiable-content rewards. Rather than only penalizing suspicious rewrites, the platform also credits rewrites that surface checkable factual substance, aligning creator incentives with answer trustworthiness. Experiments on three benchmarks show that VCR consistently achieves the largest Net defense-utility score, outperforming the strongest baseline by an average of 12.1 percentage points, and produces a win--win outcome under our empirical equivalence criterion.

cs.LG

Spatially Reconfigurable Antenna Systems for 6G: EM-based Channel Modeling, Measurements, and Orientation Design

Spatially reconfigurable antenna systems (SRASs) are recognized as a key physical-layer technology for sixth-generation (6G) systems. By dynamically adjusting each antenna element's spatial configuration, e.g., position and orientation, SRASs can revamp favorable channel conditions for reliable high-rate data transmission. However, in widely adopted channel models, antennas are typically modeled as ideal isotropic radiators, and the vectorial nature of electromagnetic (EM) propagation is neglected. This oversimplified model precludes full exploitation of the degrees of freedom offered by SRASs for performance enhancement. To address this issue, in this paper, by leveraging the theoretical framework of spherical vector wave expansion, we develop an EM-based channel model tailored for SRAS-enabled multiple-input multiple-output (MIMO) systems. The proposed EM-based channel model is applicable to antennas with arbitrary structures and intrinsically accounts for the vectorial nature of EM propagation, thereby enabling accurate characterization of EM effects such as polarization mismatch on channel gain. Full-wave simulations and experimental measurements are conducted, and the results show excellent agreement with theoretical predictions. Simulation results also reveal that antenna orientation exerts a more pronounced influence on the achievable rate than antenna displacement. Therefore, building upon the derived channel model, a manifold optimization method is proposed to maximize the sum-rate of an SRAS-enabled multiuser-MIMO system by optimizing antenna orientations. Simulation results demonstrate that the proposed scheme improves the sum-rate by up to 16.6\% and 19.3\% compared to systems employing movable antennas and conventional fixed antennas, respectively.

eess.SP

ScaleResfusion: Residual Rectified Flow based on Residual Vector Field

Real-world Image Restoration (Real-IR) aims to recover high-quality (HQ) images from complex and unknown degradations. Recent diffusion-based methods have substantially improved perceptual quality, yet two obstacles remain: methods that sample from Gaussian noise require many steps and are often less faithful to the degraded input, whereas residual-based methods that start from the low-quality (LQ) image typically train task-specific models from scratch, with optimization objectives coupled to a particular noise scheduler, and therefore cannot reuse modern pre-trained generative priors. We present \textbf{ScaleResfusion}, which rewrites residual restoration as a scheduler-independent adaptation interface for pre-trained text-to-image rectified-flow models. Its core, \textbf{Residual Rectified Flow} (RRF), inserts the residual term $R$ into the linear transport path of Rectified Flow, so that sampling starts from noisy LQ at an exact acceleration point, where the signal-to-noise ratio of the starting state is continuously controlled by the residual ratio $\gamma$. The resulting optimization target, the \textbf{residual vector field}, contains no scheduler-specific coefficients and differs from the pre-trained rectified-flow target only by the residual offset $\gamma R$; adapting a frozen billion-scale backbone therefore reduces to fitting this compact residual correction with LoRA-only training. A knowledge-distillation pipeline built around RRF further reduces sampling to as few as 4 steps. Experiments on real-world super-resolution across multiple benchmarks show that ScaleResfusion achieves state-of-the-art restoration quality and transfers consistently across pre-trained rectified-flow backbones from 2B to 9B parameters.

cs.CV

A RELHIC twin candidate near the galaxy M51

We report the discovery of a pair of H I clouds near M51 (NGC 5194) using the Five-hundred-meter Aperture Spherical radio Telescope (FAST). These clouds have no optical counterparts and are potential candidates for Reionization-Limited H I Clouds (RELHICs). We search for compact H I sources in deep FEASTS observations using SoFiA and remove objects with optical counterparts through cross-matching with the DESI Legacy Imaging Surveys. The remaining candidates are modelled as hydrostatic H I structures embedded in Navarro-Frenk-White dark matter haloes and compared with RELHIC predictions and TNG50 simulations. We identify two H I clouds, Cloud S and Cloud N, at projected distances of 70--90 kpc from M51. Each cloud has an H I mass of approximately 10^6.5 solar masses, a velocity dispersion of about 20 km/s, and no detectable optical counterpart down to a g-band surface brightness limit of approximately 27.5 mag arcsec^-2. Their stellar luminosities are constrained to be below 10^5 solar luminosities. Their H I properties are consistent with RELHIC predictions, corresponding to host halo masses of 3.7 +- 0.4 * 10^9 solar masses. Cloud S and Cloud N are promising but not definitive RELHIC candidates. A tidal origin remains possible in the interacting M51 system, especially because the clouds are unresolved by FAST and Cloud N may show a velocity gradient. Future high-resolution interferometric observations will be crucial for distinguishing between starless dark matter haloes and tidal debris.

astro-ph.GA

Ising superconductivity and anomalous metallic states in a bulk crystal with artificial unidirectional stacking layers

The two-dimensional (2D) limit in macroscopic bulk crystals provides a powerful platform for exploring exotic quantum phases. Here, we report the synthesis of a Sr0.75ClNbS2 superconductor that achieves unidirectional, parallel AA stacking-a configuration never before realized in a bulk crystal. Unlike conventional intercalation, which merely expands the interlayer spacing, our approach employs a planar Sr-Cl network to enforce a complete stacking reorganization, driving all NbS2 layers from the native antiparallel AB stacking into a unidirectional, parallel AA arrangement. This stacking switch globally breaks inversion symmetry, transforming centrosymmetric 2H-NbS2 into a noncentrosymmetric bulk crystal with D3h point group symmetry. Crucially, this structural design reproduces, in three dimensions, the electronic environment of an isolated monolayer, thereby preventing cancellation of the local Ising fields. As a result, strong Ising spin-orbit coupling and spin-split bands persist throughout the bulk. Transport measurements reveal extreme superconducting anisotropy ({\gamma} ~ 77), an in-plane upper critical field (~ 10.65 T) that far exceeds the Pauli paramagnetic limit, and clean-limit superconductivity indicative of high crystalline quality. Moreover, magnetotransport uncovers a novel magnetic-field-induced anomalous metallic state characterized by finite dissipation yet a vanishing Hall response. Direct band-structure measurements corroborate the layer-decoupled, quasi-2D electronic nature of the system. This work establishes stacking-geometry engineering as a powerful strategy to artificially enforce a globally noncentrosymmetric, quasi-2D superconducting state in bulk crystals, paving the way for designing quantum materials with tunable crystalline symmetry and electronic band topology.

cond-mat.supr-con

SPRI: SVD-Partitioned Residual Initialization for Data-Constrained MoE Upcycling

Mixture-of-Experts (MoE) models enable efficient scaling, but training them from scratch remains prohibitively expensive. MoE upcycling mitigates this cost by converting pretrained dense models into sparse MoE models. However, existing upcycling methods typically rely on large-scale continued training and often perform poorly under data-constrained supervised adaptation, due to either homogeneous experts or overly disruptive perturbations to pretrained parameters. In this setting, effective upcycling must leverage pretrained weight structure while introducing sufficient diversity among routed experts. To this end, we propose SVD-Partitioned Residual Initialization (SPRI), which distributes SVD-partitioned residuals derived from pretrained feed-forward network (FFN) weights across routed experts, introducing controlled expert diversity grounded in pretrained spectral structure. We further introduce a two-stage training strategy to improve adaptation stability. We evaluate SPRI on multilingual speech-to-text translation, where limited supervised data challenges MoE upcycling and multiple target languages provide natural routing heterogeneity. On CoVoST2 across 15 En-to-XX directions, SPRI improves average BLEU and COMET over fully fine-tuned dense models by 2.58 and 3.32 points, respectively, and outperforms the prior best MoE upcycling baseline by 3.39 BLEU and 4.34 COMET points.

cs.LG

HiFAST: An HI data calibration and imaging pipeline for the FAST IV: The stray-radiation correction

Stray radiation is a considerable challenge for radio telescopes, requiring careful assessment due to its effects. This is crucial when the strong background flux from side lobes significantly affects the total flux, especially for extended sources. In this study, we introduced the beam pattern of the L-band receiver on the Five-hundred-meter Aperture Spherical Telescope (FAST), covering various frequencies based on recent observations. We discovered that the main beam efficiency of all beams exceeds 90\% throughout the L band frequencies, with efficiency decreasing slowly as frequency increases. Subsequently, we developed a module to mitigate stray radiation effects, incorporating it into FAST's standard \HI data reduction process, referred to as \texttt{HiFAST}. Our analysis shows that side lobe flux's influence, particularly for extended sources with significant surface density gradients, necessitates detailed evaluation. Corrections for the extended M33 galaxy can reach up to 20\%. Moreover, the pattern data presented here is vital for studying HI intensity maps at high redshift. The module, along with HiFAST and beam pattern data across 15 frequency bins, can be accessed at \textrm{https://hifast.readthedocs.io}. The datasets of beam pattern presented in this paper, are openly available at \textrm{https://doi.org/10.57760/sciencedb.j00113.00266} (https://www.scidb.cn/s/bqQRNv).

astro-ph.IM

The FAST Hundred-Deg$^2$ HI Deep (HD$^2$) Survey: Early Results from the Pilot Survey

The Hundred-deg$^2$ HI Deep (HD$^2$) survey carried out with the Five-hundred-meter Aperture Spherical Telescope (FAST) is planned to map a contiguous region within the DESI DR1 footprint, achieving an effective integration time of 20 minutes for each pointing and a uniform detection sensitivity of 0.28 mJy beam$^{-1}$ at 4.8 km s$^{-1}$ resolution. We present early results from the pilot HD$^2$ survey: a 10 deg$^2$ field overlapping with HSC-SSP and the DESI EDR SV3, observed with an integration time of 7.3 minutes per beam and the rms of 0.45 mJy beam$^{-1}$ at 4.8 km s$^{-1}$ resolution. We identify 339 HI sources at $z<0.09$, corresponding to $\sim$34 detections per deg$^2$, nearly six times higher than the detection rate of the wide-field surveys. Optical counterparts are primarily identified using DESI redshifts, yielding a matching rate and correctness exceeding 90% for galaxies with $r<19.5$ mag, a substantial improvement over SDSS. Under the constraint of $r < 17.8$ mag and $0.01 < z < 0.05$, nearly 50% of galaxies in the DESI BGS samples have HI detections in this pilot survey. The optical properties of these HI-detected galaxies span nearly the entire parameter range of the DESI sample. The gas fraction scaling relations versus stellar mass, stellar mass surface density, NUV-r, and specific star formation rate are consistent with previous surveys, e.g., ALFALFA, DINGO, and xGASS. These results justify the feasibility of the full HD$^2$ survey, which will build a high-completeness HI census over a contiguous area to probe the cold gas scaling relations of galaxies over different scales.

astro-ph.GA

Entanglement Growth from Structured Initial States in Many-Body Localized Systems

Understanding how complex entanglement structures emerge is a central problem in quantum many-body physics. Recent work by Zhang et al. has considered structured initial states prepared by evolving a product state under a chaotic Hamiltonian for a finite time before quenching to the target Hamiltonian. In this setup, total entanglement entropy growth in many-body localized systems exhibits two distinct regimes, first increasing and then decreasing as the initial entanglement is tuned. In this work, we identify the physical origin of this behavior by analyzing the dynamics of both the R\'enyi entanglement entropy and the Wehrl-R\'enyi entropy in the random-field XXZ model, the latter of which characterizes multipartite entanglement. We show that a similar non-monotonic dependence on the initial entanglement also appears in the net growth of the Wehrl-R\'enyi entropy for product states polarized along the $z$-direction. The first regime is governed by a finite magnetization associated with local integrals of motion, while the second reflects inter-site correlations. In contrast, for product states in the $x/y$-direction, the entanglement growth exhibits a monotonic decay. Our results provide a more fine-grained picture of how distinct initial-state properties shape entanglement dynamics in many-body localized systems.

quant-ph

How hate spreads online and why it returns: Re-entrant phases driven by collective behavior

The 2025 Bondi Beach mass-shooting was perpetrated by individuals inspired by ISIS (Islamic State) propaganda that increasingly featured anti-Semitic hate content following the October 2023 start of the Israel-Palestine war. Similar stories hold for other types of hate attacks, e.g. against Muslims on May 18, 2026. There is an urgent need to get ahead of future threats by understanding how and when a newly created piece of hate content will spread system-wide online. We present a two-species coalescence-fragmentation model with Susceptible-Infected-Recovered dynamics that incorporates the following published empirical features: (1) New pieces of hate content tend to be generated and promoted by a subset of in-built communities on less regulated platforms. (2) These `hate' communities create links (hyperlinks) with each other and with non-hate communities across all platforms to form dynamically evolving clusters (i.e. coalescence) across which new hate content can then spread. (3) These clusters can get broken up by moderator shutdowns (i.e. fragmentation). We present numerical solutions and derive two levels of approximate mean-field theory: Effective Medium Theory (EMT) and Beyond Effective Medium Theory (BEMT). Both numerical and analytic solutions reveal that system-wide spreading is governed by re-entrant threshold phases: as the fraction of hate communities varies, the system can transition from spreading to no-spreading and back to spreading. The derived analytic formulae give explicit insight into how these phase boundaries might be manipulated to prevent system-wide spreading. More broadly, the re-entrant phase behavior warns that policies which steadily reduce the number of hate communities can initially succeed but then backfire if pushed further, suggesting that blanket requirements for platforms to simply do `more' are over-simplistic.

physics.soc-ph

A Spatially Resolved HI Survey of Seyfert Galaxies: the Role of AGN Feedback in Shaping Atomic Gas Reservoirs

Active galactic nucleus (AGN) feedback is a key ingredient in galaxy evolution, yet its impact on the cold atomic gas reservoir -- the neutral hydrogen (HI) phase -- remains poorly constrained. We present the most extensive spatially resolved HI 21-cm survey of Seyfert AGN hosts to date, based on observations with the Giant Metrewave Radio Telescope (GMRT). Our high-resolution HI maps of eight Seyfert galaxies reveal detailed kinematics and surface density distributions of their atomic gas disks. We find that AGN-host galaxies exhibit a slightly shallower HI mass-size relation than the canonical relation or the SIMBA simulation predictions; however, the measured slope remains consistent with the canonical value within $2\sigma$ uncertainties. This result suggests that AGN feedback does not significantly disrupt the global extent or large-scale structure of atomic gas reservoirs. To investigate the internal HI kinematics in greater detail, we perform a 3D kinematic forward modeling of the HI disk in UGC 4503. Our analysis reveals an elevated intrinsic velocity dispersion of $\sigma = 14.9^{+6.1}_{-3.8}$ km/s and a reduced level of rotational support, with $V/\sigma = 14.28_{-4.17}^{+4.97}$, compared to large-sample star-forming spirals. These kinematic signatures, together with localized residuals in the velocity field, indicate that AGN-driven outflows or jets may inject or indirectly affect the turbulence in the atomic gas disk, potentially regulating the cold gas reservoir. Future GMRT observations, combined with optical integral-field spectroscopy from MaNGA, will enable quantitative constraints on the role of AGN feedback in regulating star formation efficiency across a larger and more representative galaxy sample.

astro-ph.GA

TacoMAS: Test-Time Co-Evolution of Topology and Capability in LLM-based Multi-Agent Systems

Multi-agent systems (MAS) have emerged as a promising paradigm for solving complex tasks. Recent work has explored self-evolving MAS that automatically optimize agent capabilities or communication topologies. However, existing methods either learn a topology that remains fixed at inference time or adapt only the topology or capability during inference. We empirically and theoretically show that effective test-time evolution requires jointly adapting both axes, but on different time scales: capabilities should update rapidly to handle emerging subtasks, while the topology should evolve more slowly to preserve coordination stability. We then introduce TacoMAS, a test-time co-evolution framework for dynamic MAS. TacoMAS formulates MAS inference as a task of online graph adaptation, where nodes represent agents with role-specific capabilities and edges define their communication topology. During inference, a fast capability loop updates agent expertise using trajectory-level feedback, while a slow meta-LLM-driven topology loop performs agents' birth-death operations on MAS, including edge edit, agent addition, and agent removal. We further show that this fast-slow design drives MAS evolution toward a task-conditioned stable equilibrium. Experiments on four benchmarks demonstrate that TacoMAS outperforms nearly 20 multi-agent baselines, achieving an average improvement of 13.3% over the strongest baseline. The codes are released at https://github.com/chenxu2-gif/TacoMAS-MultiAgent.

cs.CL

Breaking the Trade-off: Bulk 2D Ising Superconductivity with High Tc and Giant Interlayer Spacing via a Unique Chain Intercalation in (BaS)1/3TaS2

Two-dimensional (2D) transition metal dichalcogenides (TMDs) are promising platforms for low dimensional superconductivity. However, in conventional intercalated systems, achieving a high superconducting transition temperature (Tc) often comes at the expense of reduced interlayer spacing and weakened 2D character. Here, we overcome this long-standing compromise through a unique chain-like intercalation strategy. We report the synthesis and properties of a new polymorph, (BaS)1/3TaS2, in which a distinctive Ba-S-S-Ba chain structure is inserted between TaS2 bilayers. This unique configuration breaks the bulk c axis mirror symmetry while achieving exceptional interlayer decoupling, with an inter-bilayer spacing of 12.75 {\AA}-more than three times that of pristine 2H-TaS2. By suppressing interlayer electronic coupling, this structural evolution allows local inversion symmetry breaking within individual TaS2 layers to dominate. This prevents compensation of the Ising spin-orbit fields typical of centrosymmetric bulk phases, enabling robust 2D Ising superconductivity. Remarkably, the compound exhibits an enhanced Tc without sacrificing its large interlayer spacing, thereby breaking the conventional trade-off between large spacing/high anisotropy and high Tc. Comprehensive transport, magnetic, and thermodynamic measurements confirm its robust superconducting state. Our work establishes a versatile intercalation framework for designing bulk-like 2D Ising superconductors, providing a new route to reconcile competing material demands and expanding the scope of Ising superconductivity research.

cond-mat.supr-con

The Attention Market: Interpreting Online Fair Re-ranking as Manifold Optimization under Walrasian Equilibrium

Fair re-ranking aims to promote long-tail items and enhance diversity within groups in information retrieval. While previous research on online fairness-aware re-ranking has shown promising outcomes, our comprehensive evaluation of online fair re-ranking methods over 20 settings reveals significant performance disparities among existing methods. To uncover the root causes of these inconsistencies, we reformulate fair re-ranking within an attentional market framework governed by a Walrasian Equilibrium, where the fairness is treated as a taxation cost. This market-based formulation is then coupled with manifold optimization, demonstrating that seeking this equilibrium is equivalent to performing gradient descent on a specific ranking manifold constructed by the market. Different re-ranking settings induce distinct manifold geometries, and these intrinsic geometric differences dictate the gradient landscapes and optimization trajectories. We propose ManifoldRank, an efficient online fair re-ranking algorithm. ManifoldRank adjusts gradients to align with the ranking manifold, considering various contextual settings. On the supply side, it incorporates a gradient adjustment based on different fairness requirements, accounting for associated costs. On the demand side, it empirically predicts an additional gradient adjustment term derived from the ranking scores. By integrating these two gradient adjustments, ManifoldRank effectively balances fairness and accuracy. Experimental results across multiple datasets confirm ManifoldRank's effectiveness.

cs.IR

HotComment: A Benchmark for Evaluating Popularity of Online Comments

Online comments play a crucial role in shaping public sentiment and opinion dynamics on social media. However, evaluating their popularity remains challenging, not only because it depends on linguistic quality, originality, and emotional resonance, but also because stylistic preferences vary widely across platforms and user groups, causing the same comment to resonate differently in different communities. In this work, we present HotComment, a multimodal benchmark integrating video and text modalities that comprehensively quantifies popularity from three enhanced aspects: (1) Content Quality, which evaluates semantic similarity with ground-truth human comments and extends quality assessment through four interpretable dimensions; (2) Popularity Prediction, based on trends from models trained on real-world interaction data; and (3) User Behavior Simulation, which models the distribution of platform users and approximates \textbf{engagement scores} through an agent-based framework. Furthermore, we propose StyleCmt, inspired by social ripple effects, where multiple stylistic dimensions align to amplify socially resonant expressions and suppress incongruent ones.

cs.AI