SearcharxivSearch

arXiv subjects

Xuefeng Wei

Publications and source records attributed to Xuefeng Wei.

6 recordsLinked to original sources

CArtBench: Evaluating Vision-Language Models on Chinese Art Understanding, Interpretation, and Authenticity

We introduce CARTBENCH, a museum-grounded benchmark for evaluating vision-language models (VLMs) on Chinese artworks beyond short-form recognition and QA. CARTBENCH comprises four subtasks: CURATORQA for evidence-grounded recognition and reasoning, CATALOGCAPTION for structured four-section expert-style appreciation, REINTERPRET for defensible reinterpretation with expert ratings, and CONNOISSEURPAIRS for diagnostic authenticity discrimination under visually similar confounds. CARTBENCH is built by aligning image-bearing Palace Museum objects from Wikidata with authoritative catalog pages, spanning five art categories across multiple dynasties. Across nine representative VLMs, we find that high overall CURATORQA accuracy can mask sharp drops on hard evidence linking and style-to-period inference; long-form appreciation remains far from expert references; and authenticity-oriented diagnostic discrimination stays near chance, underscoring the difficulty of connoisseur-level reasoning for current models.

cs.CL

Molecular Dynamics Simulations of Microscopic Structural Transition and Macroscopic Mechanical Properties of Magnetic Gels

Magnetic gels with embedded micro/nano-sized magnetic particles in crosslinked polymer networks can be actuated by external magnetic fields, with changes in their internal microscopic structures and macroscopic mechanical properties. We investigate the responses of such magnetic gels to an external magnetic field, by means of coarse-grained molecular dynamics simulations. We find that the dynamics of magnetic particles are determined by the interplay of between magnetic dipole-dipole interactions, polymer elasticity and thermal fluctuations. The corresponding microscopic structures formed by the magnetic particles such as elongated chains can be controlled by the external magnetic field. Furthermore, the magnetic gels can exhibit reinforced macroscopic mechanical properties, where the elastic modulus increases algebraically with the magnetic moments of the particles in the form of $\propto(m-m_{\mathrm{c}})^{2}$ when magnetic chains are formed. This simulation work can not only serve as a tool for studying the microscopic and the macroscopic responses of the magnetic gels, but also facilitate future fabrications and practical controls of magnetic composites with desired physical properties.

cond-mat.soft

FLDNet: A Foreground-Aware Network for Polyp Segmentation Leveraging Long-Distance Dependencies

Given the close association between colorectal cancer and polyps, the diagnosis and identification of colorectal polyps play a critical role in the detection and surgical intervention of colorectal cancer. In this context, the automatic detection and segmentation of polyps from various colonoscopy images has emerged as a significant problem that has attracted broad attention. Current polyp segmentation techniques face several challenges: firstly, polyps vary in size, texture, color, and pattern; secondly, the boundaries between polyps and mucosa are usually blurred, existing studies have focused on learning the local features of polyps while ignoring the long-range dependencies of the features, and also ignoring the local context and global contextual information of the combined features. To address these challenges, we propose FLDNet (Foreground-Long-Distance Network), a Transformer-based neural network that captures long-distance dependencies for accurate polyp segmentation. Specifically, the proposed model consists of three main modules: a pyramid-based Transformer encoder, a local context module, and a foreground-Aware module. Multilevel features with long-distance dependency information are first captured by the pyramid-based transformer encoder. On the high-level features, the local context module obtains the local characteristics related to the polyps by constructing different local context information. The coarse map obtained by decoding the reconstructed highest-level features guides the feature fusion process in the foreground-Aware module of the high-level features to achieve foreground enhancement of the polyps. Our proposed method, FLDNet, was evaluated using seven metrics on common datasets and demonstrated superiority over state-of-the-art methods on widely-used evaluation measures.

cs.CV

Feature Aggregation Network for Building Extraction from High-resolution Remote Sensing Images

The rapid advancement in high-resolution satellite remote sensing data acquisition, particularly those achieving submeter precision, has uncovered the potential for detailed extraction of surface architectural features. However, the diversity and complexity of surface distributions frequently lead to current methods focusing exclusively on localized information of surface features. This often results in significant intraclass variability in boundary recognition and between buildings. Therefore, the task of fine-grained extraction of surface features from high-resolution satellite imagery has emerged as a critical challenge in remote sensing image processing. In this work, we propose the Feature Aggregation Network (FANet), concentrating on extracting both global and local features, thereby enabling the refined extraction of landmark buildings from high-resolution satellite remote sensing imagery. The Pyramid Vision Transformer captures these global features, which are subsequently refined by the Feature Aggregation Module and merged into a cohesive representation by the Difference Elimination Module. In addition, to ensure a comprehensive feature map, we have incorporated the Receptive Field Block and Dual Attention Module, expanding the receptive field and intensifying attention across spatial and channel dimensions. Extensive experiments on multiple datasets have validated the outstanding capability of FANet in extracting features from high-resolution satellite images. This signifies a major breakthrough in the field of remote sensing image processing. We will release our code soon.

cs.CV

Scaling transition of active turbulence from two to three dimensions

Turbulent flows are observed in low-Reynolds active fluids. They are intrinsically different from the classical inertial turbulence and behave distinctively in two- and three-dimensions. Understanding the behaviors of this new type of turbulence and their dependence on the system dimensionality is a fundamental challenge in non-equilibrium physics. We experimentally measure flow structures and energy spectra of bacterial turbulence between two large parallel plates spaced by different heights $H$. The turbulence exhibits three regimes as H increases, resulting from the competition of bacterial length, vortex size and H. This is marked by two critical heights ($H_0$ and $H_1$) and a $H^{0.5}$ scaling law of vortex size in the large-$H$ limit. Meanwhile, the spectra display distinct universal scaling laws in quasi-two-dimensional (2D) and three-dimensional (3D) regimes, independent of bacterial activity, length and $H$, whereas scaling exponents exhibit transitions in the crossover. To understand the scaling laws, we develop a hydrodynamic model using image systems to represent the effect of no-slip confining boundaries. This model predicts universal 1 and -4 scaling on large and small length scales, respectively, and -2 and -1 on intermediate length scales in 2D and 3D, respectively, which are consistent with the experimental results. Our study suggests a framework for investigating the effect of dimensionality on non-equilibrium self-organized systems.

cond-mat.soft

Modelling Elastically-Mediated Liquid-Liquid Phase Separation

We propose a continuum theory of the liquid-liquid phase separation in an elastic network where phase-separated microscopic droplets rich in one fluid component can form as an interplay of fluids mixing, droplet nucleation, network deformation, thermodynamic fluctuation, \emph{etc}. We find that the size of the phase separated droplets decreases with the shear modulus of the elastic network in the form of $\sim[\mathrm{modulus}]^{-1/3}$ and the number density of the droplet increases almost linearly with the shear modulus $\sim[\mathrm{modulus}]$, which are verified by the experimental observations. Phase diagrams in the space of (fluid constitution, mixture interaction, network modulus) are provided, which can help to understand similar phase separations in biological cells and also to guide fabrications of synthetic cells with desired phase properties.

cond-mat.soft