SearcharxivSearch

arXiv subjects

Jing Wen

Publications and source records attributed to Jing Wen.

At least 19 recordsLinked to original sources

Framing Migration News with LLMs: Structured CoT as a Support for Human Interpretation

Frame analysis of migration news is a socially consequential task: media scholars and researchers who study how migration is narrated need tools that are not only accurate, but transparent, auditable, and accessible within the resource constraints typical of academic research groups. Existing LLM-based approaches rely on proprietary APIs and large models that raise concerns about data privacy, reproducibility and equitable access among media researchers. This work studies how a locally deployable open-source LLM can support interpretable frame analysis as an assistive tool. We introduce a Structured Chain-of-Thought (SCoT) prompting approach using Llama3-8B, enabling step-by-step justifications grounded in predefined framing categories. This structured design allows users to audit model outputs and examine alternative interpretations in a task that is inherently subjective. We evaluate our approach on a dataset of migration-related news and show that SCoT improves classification performance over zero-shot and few-shot baselines while remaining feasible on a single GPU. Then, we conduct a human-centered evaluation in which annotators assess the coherence and influence of "the model's reasoning". Results indicate that SCoT explanations are generally perceived as logical (mean score 4.1/5, though with notable variation across texts) and can prompt reflection on initial interpretations, even when disagreement persists. Our findings highlight both the potential and risks of LLM-assisted frame analysis. While structured reasoning can increase the traceability of model outputs and support critical interpretation, it can also influence human judgment in subtle ways. By enabling local deployment and emphasizing human-in-the-loop interaction, this work contributes to discussions on responsible and accessible computational tools for the study of socially impactful media narratives.

cs.CL

Evidence of triggered star formation in the Pillars of Creation from JWST observations

Stars form in molecular clouds under the influence of their local environments, yet the role of massive stellar feedback in either triggering or suppressing star formation remains a fundamental question in astrophysics. The Pillars of Creation in the Eagle Nebula, sculpted by ionizing radiation and stellar winds from massive stars in NGC 6611, offer a natural laboratory for investigating this question. Here we present high-resolution observations of the Pillars of Creation using the JWST Near Infrared Camera and Mid-Infrared Instrument, revealing 253 young stellar object (YSO) candidates. These YSO candidates show spatial correlations with the edges of feedback-driven structures, with overdensities along the boundaries. A weak trend of decreasing stellar age with increasing distance from the ionizing source was tentatively observed. There also appears to be an enhancement in the star formation rate within the past 1 Myr in this region. Such age and spatial associations suggest that while the bulk of the YSOs may have formed contemporaneously with the central cluster, a subset could be associated with triggered star formation. The JWST image of intricate structures, including a spiral-like disk and bi-reflection nebulae at the tips of Pillar I and Pillar II, further highlights the complexity of star formation processes.

astro-ph.SR

Seeing Without Eyes: 4D Human-Scene Understanding from Wearable IMUs

Understanding human activities and their surrounding environments typically relies on visual perception, yet cameras pose persistent challenges in privacy, safety, energy efficiency, and scalability. We explore an alternative: 4D perception without vision. Its goal is to reconstruct human motion and 3D scene layouts purely from everyday wearable sensors. For this we introduce IMU-to-4D, a framework that repurposes large language models for non-visual spatiotemporal understanding of human-scene dynamics. IMU-to-4D uses data from a few inertial sensors from earbuds, watches, or smartphones and predicts detailed 4D human motion together with coarse scene structure. Experiments across diverse human-scene datasets show that IMU-to-4D yields more coherent and temporally stable results than SoTA cascaded pipelines, suggesting wearable motion sensors alone can support rich 4D understanding.

cs.CV

NoPo-Avatar: Generalizable and Animatable Avatars from Sparse Inputs without Human Poses

We tackle the task of recovering an animatable 3D human avatar from a single or a sparse set of images. For this task, beyond a set of images, many prior state-of-the-art methods use accurate "ground-truth" camera poses and human poses as input to guide reconstruction at test-time. We show that pose-dependent reconstruction degrades results significantly if pose estimates are noisy. To overcome this, we introduce NoPo-Avatar, which reconstructs avatars solely from images, without any pose input. By removing the dependence of test-time reconstruction on human poses, NoPo-Avatar is not affected by noisy human pose estimates, making it more widely applicable. Experiments on challenging THuman2.0, XHuman, and HuGe100K data show that NoPo-Avatar outperforms existing baselines in practical settings (without ground-truth poses) and delivers comparable results in lab settings (with ground-truth poses).

cs.CV

Field-Trial Quantum Key Distribution with Qubit-Based Frame Synchronization

Quantum key distribution (QKD) is a cryptographic technique that uses quantum mechanical principles to enable secure key exchange. Practical deployment of QKD requires robust, cost-effective systems that can operate in challenging field environments. A major challenge is achieving reliable clock synchronization without adding hardware complexity. Conventional approaches often use separate classical light signals, which increase costs and introduce noise that degrades quantum channel performance. To address this limitation, we demonstrate a QKD system incorporating a recently proposed qubit-based distributed frame synchronization method, deployed over a metropolitan fiber network in Nanning, China. Using the polarization-encoded one-decoy-state BB84 protocol and the recently proposed qubit-based distributed frame synchronization method, our system achieves synchronization directly from the quantum signal, eliminating the need for dedicated synchronization hardware. Furthermore, to counteract dynamic polarization disturbances in urban fibers, the system integrates qubit-based polarization feedback control, enabling real-time polarization compensation through an automated polarization controller using data recovered from the qubit-based synchronization signals. During 12 hours of continuous operation, the system maintained a low average quantum bit error rate (QBER) of 1.12/%, achieving a secure key rate of 26.6 kbit/s under 18 dB channel loss. Even under a high channel loss of 40 dB, a finite-key secure rate of 115 bit/s was achieved. This study represents the first successful long-term validation of a frame-synchronization based QKD scheme in a real urban environment, demonstrating exceptional stability and high-loss tolerance, and offering an alternative for building practical, scalable, and cost-efficient quantum-secure communication networks.

quant-ph

Development of 3D Pixel Sensors via an 8-inch CMOS-Compatible Process

In the construction of High-Luminosity Large Hadron Collider (HL-LHC) and Future Circular Collider (FCC) experiments, 3D pixel sensors have become indispensable components due to their superior radiation hardness, fast response, and low power consumption. However, there are still significant challenges in the process of 3D sensors manufacturing. In this work, single devices and arrays of 3D sensors based on 30 $\mu$m epitaxial silicon wafer have been designed, simulated, fabricated, and tested. This process was developed on the 8-inch CMOS process platform of the Institute of Microelectronics of the Chinese Academy of Sciences (IMECAS). The key processes include Deep Reactive Ion Etching (DRIE) with the Bosch process, in-situ doping, and an innovative back-etching. After testing the 3D pixel sensors, we have summarized the leakage current and capacitance of devices with different sizes with respect to bias voltages. We also found that the fabricated devices were almost all successfully produced, which laid a strong foundation for subsequent large-scale mass production.

physics.ins-det

Diffusion-based Virtual Staining from Polarimetric Mueller Matrix Imaging

Polarization, as a new optical imaging tool, has been explored to assist in the diagnosis of pathology. Moreover, converting the polarimetric Mueller Matrix (MM) to standardized stained images becomes a promising approach to help pathologists interpret the results. However, existing methods for polarization-based virtual staining are still in the early stage, and the diffusion-based model, which has shown great potential in enhancing the fidelity of the generated images, has not been studied yet. In this paper, a Regulated Bridge Diffusion Model (RBDM) for polarization-based virtual staining is proposed. RBDM utilizes the bidirectional bridge diffusion process to learn the mapping from polarization images to other modalities such as H\&E and fluorescence. And to demonstrate the effectiveness of our model, we conduct the experiment on our manually collected dataset, which consists of 18,000 paired polarization, fluorescence and H\&E images, due to the unavailability of the public dataset. The experiment results show that our model greatly outperforms other benchmark methods. Our dataset and code will be released upon acceptance.

eess.IV

LIFe-GoM: Generalizable Human Rendering with Learned Iterative Feedback Over Multi-Resolution Gaussians-on-Mesh

Generalizable rendering of an animatable human avatar from sparse inputs relies on data priors and inductive biases extracted from training on large data to avoid scene-specific optimization and to enable fast reconstruction. This raises two main challenges: First, unlike iterative gradient-based adjustment in scene-specific optimization, generalizable methods must reconstruct the human shape representation in a single pass at inference time. Second, rendering is preferably computationally efficient yet of high resolution. To address both challenges we augment the recently proposed dual shape representation, which combines the benefits of a mesh and Gaussian points, in two ways. To improve reconstruction, we propose an iterative feedback update framework, which successively improves the canonical human shape representation during reconstruction. To achieve computationally efficient yet high-resolution rendering, we study a coupled-multi-resolution Gaussians-on-Mesh representation. We evaluate the proposed approach on the challenging THuman2.0, XHuman and AIST++ data. Our approach reconstructs an animatable representation from sparse inputs in less than 1s, renders views with 95.1FPS at $1024 \times 1024$, and achieves PSNR/LPIPS*/FID of 24.65/110.82/51.27 on THuman2.0, outperforming the state-of-the-art in rendering quality.

cs.CV

Timing Matters: How Using LLMs at Different Timings Influences Writers' Perceptions and Ideation Outcomes in AI-Assisted Ideation

Large Language Models (LLMs) have been widely used to support ideation in the writing process. However, whether generating ideas with the help of LLMs leads to idea fixation or idea expansion is unclear. This study examines how different timings of LLM usage - either at the beginning or after independent ideation - affect people's perceptions and ideation outcomes in a writing task. In a controlled experiment with 60 participants, we found that using LLMs from the beginning reduced the number of original ideas and lowered creative self-efficacy and self-credit, mediated by changes in autonomy and ownership. We discuss the challenges and opportunities associated with using LLMs to assist in idea generation. We propose delaying the use of LLMs to support ideation while considering users' self-efficacy, autonomy, and ownership of the ideation outcomes.

cs.HC

Mass-loss Rate of Highly Evolved Stars in the Magellanic Clouds

Asymptotic giant branch stars (AGBs) and red supergiant stars (RSGs) exhibit significant mass loss phenomena and are considered important sources of interstellar dust. In this work, we employed an uniform method of spectral energy distribution fitting to analyze a large, and hence statistically significant, sample of approximately 40,000 RSGs and AGBs in the Magellanic Clouds (MCs), providing a new catalog of evolved stars that includes stellar parameters and dust properties. Our results reveal that the total dust-production rate (DPR) of the Large Magellanic Cloud is approximately $9.69\times10^{-6}\,\rm{M_{\odot }\, yr^{-1}}$, while it is around $1.75\times10^{-6}\,\rm{M_{\odot }\,yr^{-1}}$ for the Small Magellanic Cloud, with a few stars significantly contributing to the total DPR. No significant differences were observed in the contributions to DPR from carbon-rich and oxygen-rich (O-rich) evolved stars in the MCs. We explored the relations between stellar parameters (luminosity, infrared color, period, amplitude) and mass-loss rate (MLR) for evolved stars. A prominent turning point at $\log{(L/L_{\odot})} \approx 4.4$ appears in the luminosity-MLR diagram of RSGs, potentially related to the mass-loss mechanism of RSGs. The luminosity-MLR relation of AGBs is highly scattered. The DPR of AGBs shows a clear change with pulsation period and amplitude, with DPR exhibiting a drastic increase at pulsation periods of approximately 300 days and I-band amplitudes greater than 0.5 mag. Metallicity has some impact on the DPR of O-rich stars, with lower metallicity seeming to result in lower mean DPR and a higher proportion of optically thin stars.

astro-ph.SR

AnyTaskTune: Advanced Domain-Specific Solutions through Task-Fine-Tuning

The pervasive deployment of Large Language Models-LLMs in various sectors often neglects the nuanced requirements of individuals and small organizations, who benefit more from models precisely tailored to their specific business contexts rather than those with broadly superior general capabilities. This work introduces \textbf{AnyTaskTune}, a novel fine-tuning methodology coined as \textbf{Task-Fine-Tune}, specifically developed to elevate model performance on a diverse array of domain-specific tasks. This method involves a meticulous process to identify and define targeted sub-tasks within a domain, followed by the creation of specialized enhancement datasets for fine-tuning, thereby optimizing task-specific model performance. We conducted comprehensive fine-tuning experiments not only in the legal domain for tasks such as keyword extraction and sentence prediction but across over twenty different sub-tasks derived from the domains of finance, healthcare, law, psychology, consumer services, and human resources. To substantiate our approach and facilitate community engagement, we will open-source these bilingual task datasets. Our findings demonstrate that models fine-tuned using the \textbf{Task-Fine-Tune} methodology not only achieve superior performance on these specific tasks but also significantly outperform models with higher general capabilities in their respective domains. Our work is publicly available at \url{https://github.com/PandaVT/DataTager}.

cs.CL

PuFace: Defending against Facial Cloaking Attacks for Facial Recognition Models

The recently proposed facial cloaking attacks add invisible perturbation (cloaks) to facial images to protect users from being recognized by unauthorized facial recognition models. However, we show that the "cloaks" are not robust enough and can be removed from images. This paper introduces PuFace, an image purification system leveraging the generalization ability of neural networks to diminish the impact of cloaks by pushing the cloaked images towards the manifold of natural (uncloaked) images before the training process of facial recognition models. Specifically, we devise a purifier that takes all the training images including both cloaked and natural images as input and generates the purified facial images close to the manifold where natural images lie. To meet the defense goal, we propose to train the purifier on particularly amplified cloaked images with a loss function that combines image loss and feature loss. Our empirical experiment shows PuFace can effectively defend against two state-of-the-art facial cloaking attacks and reduces the attack success rate from 69.84\% to 7.61\% on average without degrading the normal accuracy for various facial recognition models. Moreover, PuFace is a model-agnostic defense mechanism that can be applied to any facial recognition model without modifying the model structure.

cs.CV

HOLMES: to Detect Adversarial Examples with Multiple Detectors

Deep neural networks (DNNs) can easily be cheated by some imperceptible but purposeful noise added to images, and erroneously classify them. Previous defensive work mostly focused on retraining the models or detecting the noise, but has either shown limited success rates or been attacked by new adversarial examples. Instead of focusing on adversarial images or the interior of DNN models, we observed that adversarial examples generated by different algorithms can be identified based on the output of DNNs (logits). Logit can serve as an exterior feature to train detectors. Then, we propose HOLMES (Hierarchically Organized Light-weight Multiple dEtector System) to reinforce DNNs by detecting potential adversarial examples to minimize the threats they may bring in practical. HOLMES is able to distinguish \textit{unseen} adversarial examples from multiple attacks with high accuracy and low false positive rates than single detector systems even in an adaptive model. To ensure the diversity and randomness of detectors in HOLMES, we use two methods: training dedicated detectors for each label and training detectors with top-k logits. Our effective and inexpensive strategies neither modify original DNN models nor require its internal parameters. HOLMES is not only compatible with all kinds of learning models (even only with external APIs), but also complementary to other defenses to achieve higher detection rates (may also fully protect the system against various adversarial examples).

cs.AI

GoMAvatar: Efficient Animatable Human Modeling from Monocular Video Using Gaussians-on-Mesh

We introduce GoMAvatar, a novel approach for real-time, memory-efficient, high-quality animatable human modeling. GoMAvatar takes as input a single monocular video to create a digital avatar capable of re-articulation in new poses and real-time rendering from novel viewpoints, while seamlessly integrating with rasterization-based graphics pipelines. Central to our method is the Gaussians-on-Mesh representation, a hybrid 3D model combining rendering quality and speed of Gaussian splatting with geometry modeling and compatibility of deformable meshes. We assess GoMAvatar on ZJU-MoCap data and various YouTube videos. GoMAvatar matches or surpasses current monocular human modeling algorithms in rendering quality and significantly outperforms them in computational efficiency (43 FPS) while being memory-efficient (3.63 MB per subject).

cs.CV

Evolved Massive Stars at Low-metallicity VII. the Lower Mass Limit of Red Supergiant Population in the Large Magellanic Cloud

The precise definition of the lower mass limit of red supergiant stars (RSGs) is an open question in astrophysics and does not attract too much attention. Here we assemble a spectroscopic evolved cool star sample with 6,602 targets, including RSGs, asymptotic giant branch stars, and red giant branch stars, in the Large Magellanic Cloud based on \textit{Gaia} DR3 and SDSS-IV/APOGEE-2. The reference spectrum of each stellar population is built according to the quantile range of relative intensity ($1\%\sim99\%$). Five different methods, e.g., chi-square ($\chi^2$), cosine similarity (CS), machine learning (ML), equivalent width (EW), and line ratio (LR), are used in order to separate different stellar populations. The ML and $\chi^2$ provide the best and relatively consistent prediction of certain population. The derived lower limit of the RSG population is able to reach to the $\rm K_S$-band tip of red giant branch ($\rm K_S~$$\approx12.0$ mag), indicating a luminosity as low as about $10^{3.5}~L_{\sun}$, which corresponds to a stellar radius only about $100~R_{\sun}$. Given the mass-luminosity relation of $L/L_\sun =f(M/M_\sun)^3$ with $f\approx15.5\pm3$ and taking into account of the mass loss of faint RSGs up to now, the minimal initial mass of the RSG population would be about $6.1\pm0.4~M_\sun$, which is much lower than the traditional threshold of $8~M_\sun$ for the massive stars. This is the first spectroscopic evidence, indicating that the lower mass limit of RSG population is around $6~M_\sun$. However, the destinies of such faint RSGs are still elusive and may have large impact on the stellar evolutionary and supernova models.

astro-ph.SR

Evolved Massive Stars at Low-metallicity VI. Mass-Loss Rate of Red Supergiant Stars in the Large Magellanic Cloud

Mass loss is a crucial process that affects the observational properties, evolution path and fate of highly evolved stars. However, the mechanism of mass loss is still unclear, and the mass-loss rate (MLR) of red supergiant stars (RSGs) requires further research and precise evaluation. To address this, we utilized an updated and complete sample of RSGs in the Large Magellanic Cloud (LMC) and employed the 2-DUST radiation transfer model and spectral energy distribution (SED) fitting approach to determine the dust-production rates (DPRs) and dust properties of the RSGs. We have fitted 4,714 selected RSGs with over 100,000 theoretical templates of evolved stars. Our results show that the DPR range of RSGs in the LMC is $10^{-11}\, \rm{M_{\odot}\, yr^{-1}}$ to $10^{-7}\, \rm{M_{\odot}\, yr^{-1}}$, and the total DPR of all RSGs is 1.14 $\times 10^{-6} \, \rm{M_{\odot} \, yr^{-1}}$. We find that $63.3\%$ RSGs are oxygen-rich, and they account for $97.2\%$ of the total DPR. The optically thin RSG, which comprise $30.6\%$ of our sample, contribute only $0.1\%$ of the total DPR, while carbon-rich RSGs ($6.1\%$) produce $2.7\%$ of the total DPR. Overall, 208 RSGs contributed $76.6\%$ of the total DPR. We have established a new relationship between the MLR and luminosity of RSGs in the LMC, which exhibits a positive trend and a clear turning point at $\log{L/L_{\odot}} \approx 4.4$.

astro-ph.GA

Machine Mindset: An MBTI Exploration of Large Language Models

We present a novel approach for integrating Myers-Briggs Type Indicator (MBTI) personality traits into large language models (LLMs), addressing the challenges of personality consistency in personalized AI. Our method, "Machine Mindset," involves a two-phase fine-tuning and Direct Preference Optimization (DPO) to embed MBTI traits into LLMs. This approach ensures that models internalize these traits, offering a stable and consistent personality profile. We demonstrate the effectiveness of our models across various domains, showing alignment between model performance and their respective MBTI traits. The paper highlights significant contributions in the development of personality datasets and a new training methodology for personality integration in LLMs, enhancing the potential for personalized AI applications. We also open-sourced our model and part of the data at \url{https://github.com/PKU-YuanGroup/Machine-Mindset}.

cs.CL

Evolved Massive Stars at Low-metallicity V. Mass-Loss Rate of Red Supergiant Stars in the Small Magellanic Cloud

We assemble the most complete and clean red supergiant (RSG) sample (2,121 targets) so far in the Small Magellanic Cloud (SMC) with 53 different bands of data to study the MLR of RSGs. In order to match the observed spectral energy distributions (SEDs), a theoretical grid of 17,820 Oxygen-rich models (``normal'' and ``dusty'' grids are half-and-half) is created by the radiatively-driven wind model of the DUSTY code, covering a wide range of dust parameters. We select the best model for each target by calculating the minimal modified chi-square and visual inspection. The resulting MLRs from DUSTY are converted to real MLRs based on the scaling relation, for which a total MLR of $6.16\times10^{-3}$ $M_\odot$ yr$^{-1}$ is measured (corresponding to a dust-production rate of $\sim6\times10^{-6}$ $M_\odot$ yr$^{-1}$), with a typical MLR of $\sim10^{-6}$ $M_\odot$ yr$^{-1}$ for the general population of the RSGs. The complexity of mass-loss estimation based on the SED is fully discussed for the first time, indicating large uncertainties based on the photometric data (potentially up to one order of magnitude or more). The Hertzsprung-Russell and luminosity versus median absolute deviation diagrams of the sample indicate the positive relation between luminosity and MLR. Meanwhile, the luminosity versus MLR diagrams show a ``knee-like'' shape with enhanced mass-loss occurring above $\log_{10}(L/L_\odot)\approx4.6$, which may be due to the degeneracy of luminosity, pulsation, low surface gravity, convection, and other factors. We derive our MLR relation by using a third-order polynomial to fit the sample and compare our result with previous empirical MLR prescriptions. Given that our MLR prescription is based on a much larger sample than previous determinations, it provides a more accurate relation at the cool and luminous region of the H-R diagram at low-metallicity compared to previous studies.

astro-ph.SR