Searcharxiv⌕ Search

arXiv subjects

Ping Li

Publications and source records attributed to Ping Li.

At least 73 records · Page 4Linked to original sources

Constructing disjoint Steiner trees in Sierpiński graphs

Let $G$ be a graph and $S\subseteq V(G)$ with $|S|\geq 2$. Then the trees $T_1, T_2, \cdots, T_\ell$ in $G$ are \emph{internally disjoint Steiner trees} connecting $S$ (or $S$-Steiner trees) if $E(T_i) \cap E(T_j )=\emptyset$ and $V(T_i)\cap V(T_j)=S$ for every pair of distinct integers $i,j$, $1 \leq i, j \leq \ell$. Similarly, if we only have the condition $E(T_i) \cap E(T_j )=\emptyset$ but without the condition $V(T_i)\cap V(T_j)=S$, then they are \emph{edge-disjoint Steiner trees}. The \emph{generalized $k$-connectivity}, denoted by $κ_k(G)$, of a graph $G$, is defined as $κ_k(G)=\min\{κ_G(S)|S \subseteq V(G) \ \textrm{and} \ |S|=k \}$, where $κ_G(S)$ is the maximum number of internally disjoint $S$-Steiner trees. The \emph{generalized local edge-connectivity} $λ_{G}(S)$ is the maximum number of edge-disjoint Steiner trees connecting $S$ in $G$. The {\it generalized $k$-edge-connectivity} $λ_k(G)$ of $G$ is defined as $λ_k(G)=\min\{λ_{G}(S)\,|\,S\subseteq V(G) \ and \ |S|=k\}$. These measures are generalizations of the concepts of connectivity and edge-connectivity, and they and can be used as measures of vulnerability of networks. It is, in general, difficult to compute these generalized connectivities. However, there are precise results for some special classes of graphs. In this paper, we obtain the exact value of $λ_{k}(S(n,\ell))$ for $3\leq k\leq \ell^n$, and the exact value of $κ_{k}(S(n,\ell))$ for $3\leq k\leq \ell$, where $S(n, \ell)$ is the Sierpiński graphs with order $\ell^n$. As a direct consequence, these graphs provide additional interesting examples when $λ_{k}(S(n,\ell))=κ_{k}(S(n,\ell))$. We also study the some network properties of Sierpiński graphs.

math.CO↗

Beware of Calibration Data for Pruning Large Language Models

As large language models (LLMs) are widely applied across various fields, model compression has become increasingly crucial for reducing costs and improving inference efficiency. Post-training pruning is a promising method that does not require resource-intensive iterative training and only needs a small amount of calibration data to assess the importance of parameters. Recent research has enhanced post-training pruning from different aspects but few of them systematically explore the effects of calibration data, and it is unclear if there exist better calibration data construction strategies. We fill this blank and surprisingly observe that calibration data is also crucial to post-training pruning, especially for high sparsity. Through controlled experiments on important influence factors of calibration data, including the pruning settings, the amount of data, and its similarity with pre-training data, we observe that a small size of data is adequate, and more similar data to its pre-training stage can yield better performance. As pre-training data is usually inaccessible for advanced LLMs, we further provide a self-generating calibration data synthesis strategy to construct feasible calibration data. Experimental results on recent strong open-source LLMs (e.g., DCLM, and LLaMA-3) show that the proposed strategy can enhance the performance of strong pruning methods (e.g., Wanda, DSnoT, OWL) by a large margin (up to $2.68\%$). Code is available at https://github.com/Dereck0602/calibration_data.

cs.CL↗

Improvement on LiDAR-Camera Calibration Using Square Targets

Precise sensor calibration is critical for autonomous vehicles as a prerequisite for perception algorithms to function properly. Rotation error of one degree can translate to position error of meters in target object detection at large distance, leading to improper reaction of the system or even safety related issues. Many methods for multi-sensor calibration have been proposed. However, there are very few work that comprehensively consider the challenges of the calibration procedure when applied to factory manufacturing pipeline or after-sales service scenarios. In this work, we introduce a fully automatic LiDAR-camera extrinsic calibration algorithm based on targets that is fast, easy to deploy and robust to sensor noises such as missing data. The core of the method include: (1) an automatic multi-stage LiDAR board detection pipeline using only geometry information with no specific material requirement; (2) a fast coarse extrinsic parameter search mechanism that is robust to initial extrinsic errors; (3) a direct optimization algorithm that is robust to sensor noises. We validate the effectiveness of our methods through experiments on data captured in real world scenarios.

cs.RO↗

Alternating Regret for Online Convex Optimization

Motivated by alternating learning dynamics in two-player games, a recent work by Cevher et al.(2024) shows that $o(\sqrt{T})$ alternating regret is possible for any $T$-round adversarial Online Linear Optimization (OLO) problem, and left as an open question whether the same is true for general Online Convex Optimization (OCO). We answer this question in the affirmative by showing that the continuous Hedge algorithm achieves $\tilde{\mathcal{O}}(d^{\frac{2}{3}}T^{\frac{1}{3}})$ alternating regret for any adversarial $d$-dimensional OCO problems. We show that this implies an alternating learning dynamic that finds a Nash equilibrium for any convex-concave zero-sum games or a coarse correlated equilibrium for any convex two-player general-sum games at a rate of $\tilde{\mathcal{O}}(d^{\frac{2}{3}}/T^{\frac{2}{3}})$. To further improve the time complexity and/or the dimension dependence, we propose another simple algorithm, Follow-the-Regularized-Leader with a regularizer whose convex conjugate is 3rd-order smooth, for OCO with smooth and self-concordant loss functions (such as linear or quadratic losses). We instantiate our algorithm with different regularizers and show that, for example, when the decision set is the $\ell_2$ ball, our algorithm achieves $\tilde{\mathcal{O}}(T^{\frac{2}{5}})$ alternating regret with no dimension dependence (and a better $\tilde{\mathcal{O}}(T^{\frac{1}{3}})$ bound for quadratic losses). We complement our results by showing some algorithm-specific alternating regret lower bounds, including a somewhat surprising $Ω(\sqrt{T})$ lower bound for a Regret Matching variant that is widely used in alternating learning dynamics.

cs.LG↗

Efficient Star Distillation Attention Network for Lightweight Image Super-Resolution

In recent years, the performance of lightweight Single-Image Super-Resolution (SISR) has been improved significantly with the application of Convolutional Neural Networks (CNNs) and Large Kernel Attention (LKA). However, existing information distillation modules for lightweight SISR struggle to map inputs into High-Dimensional Non-Linear (HDNL) feature spaces, limiting their representation learning. And their LKA modules possess restricted ability to capture the multi-shape multi-scale information for long-range dependencies while encountering a quadratic increase in the computational burden with increasing convolutional kernel size of its depth-wise convolutional layer. To address these issues, we firstly propose a Star Distillation Module (SDM) to enhance the discriminative representation learning via information distillation in the HDNL feature spaces. Besides, we present a Multi-shape Multi-scale Large Kernel Attention (MM-LKA) module to learn representative long-range dependencies while incurring low computational and memory footprints, leading to improving the performance of CNN-based self-attention significantly. Integrating SDM and MM-LKA, we develop a Residual Star Distillation Attention Module (RSDAM) and take it as the building block of the proposed efficient Star Distillation Attention Network (SDAN) which possesses high reconstruction efficiency to recover a higher-quality image from the corresponding low-resolution (LR) counterpart. When compared with other lightweight state-of-the-art SISR methods, extensive experiments show that our SDAN with low model complexity yields superior performance quantitatively and visually.

eess.IV↗

Type III Valley Polarization and Anomalous Valley Hall Effect in Two-Dimensional Non-Janus and Janus Altermagnet Fe2WS2Se2

Exploiting the valley degree of freedom introduces a novel paradigm for advancing quantum information technology. Currently, the investigation on spontaneous valley polarization mainly focuses on two major types of systems. One type magnetic systems by breaking the time-reversal symmetry, the other is ferroelectric materials through breaking the inversion symmetry. Might there be additional scenarios? Here, we propose to realize spontaneous valley polarization by breaking the mirror symmetry in the altermagnets, named type III valley polarization. Through symmetry analysis and first-principles calculations, we confirm that this mechanism is feasible in Non-Janus Fe2WS2Se2. Monolayer Non-Janus and Janus Fe2WS2Se2 are stable Neel-type antiferromagnetic state with the direct band gap semiconductor. More interestingly, their magnetic anisotropy energy exhibits the rare biaxial anisotropy and a four-leaf clover shape in the xy plane, while the xz and yz planes show the common uniaxial anisotropy. This originated from the fourth-order single ion interactions. More importantly, the valley splitting is spontaneously generated in the Non-Janus Fe2WS2Se2 due to the Mxy symmetry breaking, without requiring the SOC effect. Both the Non-Janus and Janus Fe2WS2Se2 exhibit diverse valley polarization and anomalous valley Hall effect properties. In addition, the magnitude and direction of valley polarization can be effectively tuned by the biaxial strain and magnetic field. Our findings not only expand the realization system of spontaneous valley polarization, but also provide a theoretical basis for the high-density storage of valley degrees of freedom.

cond-mat.mtrl-sci↗

AI shares emotion with humans across languages and cultures

Effective and safe human-machine collaboration requires the regulated and meaningful exchange of emotions between humans and artificial intelligence (AI). Current AI systems based on large language models (LLMs) can provide feedback that makes people feel heard. Yet it remains unclear whether LLMs represent emotion in language as humans do, or whether and how the emotional tone of their output can be controlled. We assess human-AI emotional alignment across linguistic-cultural groups and model-families, using interpretable LLM features translated from concept-sets for over twenty nuanced emotion categories (including six basic emotions). Our analyses reveal that LLM-derived emotion spaces are structurally congruent with human perception, underpinned by the fundamental affective dimensions of valence and arousal. Furthermore, these emotion-related features also accurately predict large-scale behavioural data on word ratings along these two core dimensions, reflecting both universal and language-specific patterns. Finally, by leveraging steering vectors derived solely from human-centric emotion concepts, we show that model expressions can be stably and naturally modulated across distinct emotion categories, which provides causal evidence that human emotion concepts can be used to systematically induce LLMs to produce corresponding affective states when conveying content. These findings suggest AI not only shares emotional representations with humans but its affective outputs can be precisely guided using psychologically grounded emotion concepts.

cs.CL↗

V455 Car: an oscillating eclipsing Algol-type binary in triple star system

V455 Car is a southern oscillating eclipsing Algol-type system with an orbital period of 5.132888 days. Our first photometric solutions based on the Transiting Exoplanet Survey Satellite indicate that it is a semi-detached binary with the secondary star is almost filling its Roche lobe. The noticeable O'Connell effect in light curve could be explained by hot spot on the primary component, which may be attributed to the mass transfer from the secondary component to the primary one. The absolute parameters are determined as: $M_{1} = 5.30 \pm 1.10 \, \rm M_{\odot}$, $R_{1} = 3.17 \pm 0.22 \, \rm R_{\odot}$ for the primary, and $M_{2} = 1.58 \pm 0.32 \, \rm M_{\odot}$, $R_{2} = 6.66 \pm 0.46 \, \rm R_{\odot}$ for the secondary. \textbf{Based on $O-C$ analysis, we find a periodic variation of $P_3=26.62(\pm1.66)\,yr$. The periodic oscillation suggests a possible third body with a minimal mass of $0.59(\pm0.13)\,\rm M_{\odot}$}. It is speculated that the secondary star has undergone a longer evolution, leading to a mass ratio reversal being experienced in the binary system. Our frequency analysis finds that the primary of V455 Car may be an SPB/SLF star. This study reports a novel example of an oscillating eclipsing Algol-type system featuring an SPB/SLF primary star and a red giant star, which suggest that strong observational results for a high incidence of third bodies in massive binaries.

astro-ph.SR↗

Beyond Templates: Dynamic Adaptation of Reasoning Demonstrations via Feasibility-Aware Exploration

Large language models (LLMs) have shown remarkable reasoning capabilities, yet aligning such abilities to small language models (SLMs) remains a challenge due to distributional mismatches and limited model capacity. Existing reasoning datasets, typically designed for powerful LLMs, often lead to degraded performance when directly applied to weaker models. In this work, we introduce Dynamic Adaptation of Reasoning Trajectories (DART), a novel data adaptation framework that bridges the capability gap between expert reasoning trajectories and diverse SLMs. Instead of uniformly imitating expert steps, DART employs a selective imitation strategy guided by step-wise adaptability estimation via solution simulation. When expert steps surpass the student's capacity -- signaled by an Imitation Gap -- the student autonomously explores alternative reasoning paths, constrained by outcome consistency. We validate DART across multiple reasoning benchmarks and model scales, demonstrating that it significantly improves generalization and data efficiency over static fine-tuning. Our method enhances supervision quality by aligning training signals with the student's reasoning capabilities, offering a scalable solution for reasoning alignment in resource-constrained models.

cs.CL↗

Temporal Consistency Constrained Transferable Adversarial Attacks with Background Mixup for Action Recognition

Action recognition models using deep learning are vulnerable to adversarial examples, which are transferable across other models trained on the same data modality. Existing transferable attack methods face two major challenges: 1) they heavily rely on the assumption that the decision boundaries of the surrogate (a.k.a., source) model and the target model are similar, which limits the adversarial transferability; and 2) their decision boundary difference makes the attack direction uncertain, which may result in the gradient oscillation, weakening the adversarial attack. This motivates us to propose a Background Mixup-induced Temporal Consistency (BMTC) attack method for action recognition. From the input transformation perspective, we design a model-agnostic background adversarial mixup module to reduce the surrogate-target model dependency. In particular, we randomly sample one video from each category and make its background frame, while selecting the background frame with the top attack ability for mixup with the clean frame by reinforcement learning. Moreover, to ensure an explicit attack direction, we leverage the background category as guidance for updating the gradient of adversarial example, and design a temporal gradient consistency loss, which strengthens the stability of the attack direction on subsequent frames. Empirical studies on two video datasets, i.e., UCF101 and Kinetics-400, and one image dataset, i.e., ImageNet, demonstrate that our method significantly boosts the transferability of adversarial examples across several action/image recognition models. Our code is available at https://github.com/mlvccn/BMTC_TransferAttackVid.

cs.CV↗

Accurate KV Cache Quantization with Outlier Tokens Tracing

The impressive capabilities of Large Language Models (LLMs) come at the cost of substantial computational resources during deployment. While KV Cache can significantly reduce recomputation during inference, it also introduces additional memory overhead. KV Cache quantization presents a promising solution, striking a good balance between memory usage and accuracy. Previous research has shown that the Keys are distributed by channel, while the Values are distributed by token. Consequently, the common practice is to apply channel-wise quantization to the Keys and token-wise quantization to the Values. However, our further investigation reveals that a small subset of unusual tokens exhibit unique characteristics that deviate from this pattern, which can substantially impact quantization accuracy. To address this, we develop a simple yet effective method to identify these tokens accurately during the decoding process and exclude them from quantization as outlier tokens, significantly improving overall accuracy. Extensive experiments show that our method achieves significant accuracy improvements under 2-bit quantization and can deliver a 6.4 times reduction in memory usage and a 2.3 times increase in throughput.

cs.CL↗

SIR: Multi-view Inverse Rendering with Decomposable Shadow Under Indoor Intense Lighting

We propose SIR, an efficient method to decompose differentiable shadows for inverse rendering on indoor scenes using multi-view data, addressing the challenges in accurately decomposing the materials and lighting conditions. Unlike previous methods that struggle with shadow fidelity in complex lighting environments, our approach explicitly learns shadows for enhanced realism in material estimation under unknown light positions. Utilizing posed HDR images as input, SIR employs an SDF-based neural radiance field for comprehensive scene representation. Then, SIR integrates a shadow term with a three-stage material estimation approach to improve SVBRDF quality. Specifically, SIR is designed to learn a differentiable shadow, complemented by BRDF regularization, to optimize inverse rendering accuracy. Extensive experiments on both synthetic and real-world indoor scenes demonstrate the superior performance of SIR over existing methods in both quantitative metrics and qualitative analysis. The significant decomposing ability of SIR enables sophisticated editing capabilities like free-view relighting, object insertion, and material replacement. The code and data are available at https://xiaokangwei.github.io/SIR/.

cs.CV↗

Influence of molecular rotation on the generation of N$_2^+$ air lasing

N$_2^+$ air lasing has attracted considerable attention due to its promising applications in remote sensing and the debates surrounding its generation mechanisms. Here, we present a comprehensive theoretical investigation of the role of molecular rotation in N$_2^+$ lasing at 391 nm ($B^2 Σ_u^+(v''=0)\rightarrow X^2 Σ_g^+ (v=0)$). By solving the open-system density matrix and Maxwell-Bloch equations in a rovibronic-state basis, we examine both the formation of the N$_2^+$ gain medium induced by a femtosecond pump pulse and the subsequent spatial propagation of the seed pulse. During the pump stage, rotational dynamics are found to significantly modify the angle-dependent populations of ionic vibrational-electronic states within tens of femtoseconds. Furthermore, ionization-produced rotational coherences substantially enhance the population inversion between the $X^2 Σ_g^+ (v=0)$ and $B^2 Σ_u^+(v''=0)$ states. In the seed propagation stage, both population inversion and rotational coherence are found to contribute to the lasing process, with the latter playing a dominant role in amplifying the lasing signals. These findings reveal the crucial role of molecular rotation in N$_2^+$ air lasing and highlight its potential as a tunable parameter for controlling lasing dynamics.

physics.optics↗

Valley Polarization and Anomalous Valley Hall Effect in Altermagnet Ti2Se2S with Multipiezo Properties

Recently, altermagnets demonstrate numerous newfangle physical phenomena due to their inherent antiferromagnetic coupling and spontaneous spin splitting, that are anticipated to enable innovative spintronic devices. However, the rare two-dimensional altermagnets have been reported, making it difficult to meet the requirements for high-performance spintronic devices on account of the growth big data. Here, we predict a stable monolayer Ti2Se2S with out-of-plane altermagnetic ground state and giant valley splitting. The electronic properties of altermagnet Ti2Se2S are highly dependent on the onsite electron correlation. Through symmetry analysis, we find that the valleys of X and Y points are protected by the mirror Mxy symmetry rather than the time-reversal symmetry. Therefore, the multipiezo effect, including piezovalley and piezomagnetism, can be induced by the uniaxial strain. The total valley splitting of monolayer Ti2Se2S can be as high as ~500 meV. Most interestingly, the direction of valley polarization can be effectively tuned by the uniaxial strain, based on this, we have defined logical "0", "+1", and "-1" states for data transmission and storage. In addition, we have designed a schematic diagram for observing the anomalous Hall effect in experimentally. Our findings have enriched the candidate materials of two-dimensional altermagnet for the ultra-fast and low power consumption device applications.

cond-mat.mtrl-sci↗

NoiseController: Towards Consistent Multi-view Video Generation via Noise Decomposition and Collaboration

High-quality video generation is crucial for many fields, including the film industry and autonomous driving. However, generating videos with spatiotemporal consistencies remains challenging. Current methods typically utilize attention mechanisms or modify noise to achieve consistent videos, neglecting global spatiotemporal information that could help ensure spatial and temporal consistency during video generation. In this paper, we propose the NoiseController, consisting of Multi-Level Noise Decomposition, Multi-Frame Noise Collaboration, and Joint Denoising, to enhance spatiotemporal consistencies in video generation. In multi-level noise decomposition, we first decompose initial noises into scene-level foreground/background noises, capturing distinct motion properties to model multi-view foreground/background variations. Furthermore, each scene-level noise is further decomposed into individual-level shared and residual components. The shared noise preserves consistency, while the residual component maintains diversity. In multi-frame noise collaboration, we introduce an inter-view spatiotemporal collaboration matrix and an intra-view impact collaboration matrix , which captures mutual cross-view effects and historical cross-frame impacts to enhance video quality. The joint denoising contains two parallel denoising U-Nets to remove each scene-level noise, mutually enhancing video generation. We evaluate our NoiseController on public datasets focusing on video generation and downstream tasks, demonstrating its state-of-the-art performance.

cs.CV↗

Dense $2$-connected planar graphs and the planar Turán number of $2C_k$

Shi, Walsh and Yu demonstrated that any dense planar graph with certain property (known as circuit graph) contains a large near-triangulation. We extend the result to $2$-connected plane graphs, thereby addressing a question posed by them. Using the result, we prove that the planar Tuán number of $2C_k$ is $\left[3-Θ(k^{\log_23})^{-1}\right]n$ when $k\geq 5$.

math.CO↗

Sample-level Adaptive Knowledge Distillation for Action Recognition

Knowledge Distillation (KD) compresses neural networks by learning a small network (student) via transferring knowledge from a pre-trained large network (teacher). Many endeavours have been devoted to the image domain, while few works focus on video analysis which desires training much larger model making it be hardly deployed in resource-limited devices. However, traditional methods neglect two important problems, i.e., 1) Since the capacity gap between the teacher and the student exists, some knowledge w.r.t. difficult-to-transfer samples cannot be correctly transferred, or even badly affects the final performance of student, and 2) As training progresses, difficult-to-transfer samples may become easier to learn, and vice versa. To alleviate the two problems, we propose a Sample-level Adaptive Knowledge Distillation (SAKD) framework for action recognition. In particular, it mainly consists of the sample distillation difficulty evaluation module and the sample adaptive distillation module. The former applies the temporal interruption to frames, i.e., randomly dropout or shuffle the frames during training, which increases the learning difficulty of samples during distillation, so as to better discriminate their distillation difficulty. The latter module adaptively adjusts distillation ratio at sample level, such that KD loss dominates the training with easy-to-transfer samples while vanilla loss dominates that with difficult-to-transfer samples. More importantly, we only select those samples with both low distillation difficulty and high diversity to train the student model for reducing computational cost. Experimental results on two video benchmarks and one image benchmark demonstrate the superiority of the proposed method by striking a good balance between performance and efficiency.

cs.CV↗

The maximum number of cliques in disjoint copies of graphs

The problem of determining the maximum number of copies of $T$ in an $H$-free graph, for any graphs $T$ and $H$, was considered by Alon and Shikhelman. This is a variant of Turán's classical extremal problem. We show lower and upper bounds for the maximum number of $s$-cliques in a graph with no disjoint copies of arbitrary graph. We also determine the maximum number of $s$-cliques in an $n$-vertex graph that does not contain a disjoint union of $k$ paths of length two when $k=2,3$, or $s\geqslant k+2$, or $n$ is sufficiently large, this partly confirms a conjecture posed by Chen, Yang, Yuan, and Zhang \cite{2024Chen113974}.

math.CO↗