SearcharxivSearch

arXiv subjects

Yuxi Guo

Publications and source records attributed to Yuxi Guo.

5 recordsLinked to original sources

Greedy-Gnorm: A Gradient Matrix Norm-Based Alternative to Attention Entropy for Head Pruning

Attention head pruning has emerged as an effective technique for transformer model compression, an increasingly important goal in the era of Green AI. However, existing pruning methods often rely on static importance scores, which fail to capture the evolving role of attention heads during iterative removal. We propose Greedy-Gradient norm (Greedy-Gnorm), a novel head pruning algorithm that dynamically recalculates head importance after each pruning step. Specifically, each head is scored by the elementwise product of the l2-norms of its Q/K/V gradient blocks, as estimated from a hold-out validation set and updated at every greedy iteration. This dynamic approach to scoring mitigates against stale rankings and better reflects gradient-informed importance as pruning progresses. Extensive experiments on BERT, ALBERT, RoBERTa, and XLM-RoBERTa demonstrate that Greedy-Gnorm consistently preserves accuracy under substantial head removal, outperforming attention entropy. By effectively reducing model size while maintaining task performance, Greedy-Gnorm offers a promising step toward more energy-efficient transformer model deployment.

cs.LG

Central Limit Theorem for Irregular Discretization Scheme of Multilevel Monte Carlo Method

In this paper, we study the asymptotic error distribution for a two-level irregular discretization scheme of the solution to the stochastic differential equations (SDE for short) driven by a continuous semimartingale and obtain a central limit theorem for the error processes with the rate $\sqrt{n}$. As an application, in the spirit of the result of Ben Alaya and Kebaier, we get a central limit theorem of the Linderberg-Feller type for the irregular discretization scheme of the multilevel Monte Carlo method.

math.PR

Physics-informed Neural Networks Enable High Fidelity Shear Wave Viscoelastography across Multiple organs

Tissue viscoelasticity has been recognized as a crucial biomechanical indicator for disease diagnosis and therapeutic monitoring. Conventional shear wave elastography techniques depend on dispersion analysis and face fundamental limitations in clinical scenarios. Particularly, limited wave propagation data with low signal-to-noise ratios, along with challenges in discriminating between dual dispersion sources stemming from viscoelasticity and finite tissue dimensions, pose great difficulties for extracting dispersion relation. In this study, we introduce SWVE-Net, a framework for shear wave viscoelasticity imaging based on a physics-informed neural network (PINN). SWVE-Net circumvents dispersion analysis by directly incorporating the viscoelasticity wave motion equation into the loss functions of the PINN. Finite element simulations reveal that SWVE-Net quantifies viscosity parameters within a wide range (0.15-1.5 Pa*s), even for samples just a few millimeters in size, where substantial wave reflections and dispersion occur. Ex vivo experiments demonstrate its applicability across various organs, including brain, liver, kidney, and spleen, each with distinct viscoelasticity. In in vivo human trials on breast and skeletal muscle tissues, SWVE-Net reliably assesses viscoelastic properties with standard deviation-to-mean ratios below 15%, highlighting robustness under real-world constraints. SWVE-Net overcomes the core limitations of conventional elastography and enables reliable viscoelastic characterization where traditional methods fall short. It holds promise for applications such as grading hepatic lipid accumulation, detecting myocardial infarction boundaries, and distinguishing malignant from benign tumors.

physics.med-ph

Electrical Impedance Tomography Based Closed-loop Tumor Treating Fields in Dynamic Lung Tumors

Tumor Treating Fields (TTFields) is a non-invasive anticancer modality that utilizes alternating electric fields to disrupt cancer cell division and growth. While generally well-tolerated with minimal side effects, traditional TTFields therapy for lung tumors faces challenges due to the influence of respiratory motion. We design a novel closed-loop TTFields strategy for lung tumors by incorporating electrical impedance tomography (EIT) for real-time respiratory phase monitoring and dynamic parameter adjustments. Furthermore, we conduct theoretical analysis to evaluate the performance of the proposed method using the lung motion model. Compared to conventional TTFields settings, we observed that variations in the electrical conductivity of lung during different respiratory phases led to a decrease in the average electric field intensity within lung tumors, transitioning from end-expiratory (1.08 V/cm) to end-inspiratory (0.87 V/cm) phases. Utilizing our proposed closed-Loop TTFields approach at the same dose setting (2400 mA, consistent with the traditional TTFields setting), we can achieve a higher and consistent average electric field strength at the tumor site (1.30 V/cm) across different respiratory stages. Our proposed closed-loop TTFields method has the potential to improved lung tumor therapy by mitigating the impact of respiratory motion.

physics.med-ph

Machine learning driven synthesis of few-layered WTe2

Reducing the lateral scale of two-dimensional (2D) materials to one-dimensional (1D) has attracted substantial research interest not only to achieve competitive electronic device applications but also for the exploration of fundamental physical properties. Controllable synthesis of high-quality 1D nanoribbons (NRs) is thus highly desirable and essential for the further study. Traditional exploration of the optimal synthesis conditions of novel materials is based on the trial-and-error approach, which is time consuming, costly and laborious. Recently, machine learning (ML) has demonstrated promising capability in guiding material synthesis through effectively learning from the past data and then making recommendations. Here, we report the implementation of supervised ML for the chemical vapor deposition (CVD) synthesis of high-quality 1D few-layered WTe2 nanoribbons (NRs). The synthesis parameters of the WTe2 NRs are optimized by the trained ML model. On top of that, the growth mechanism of as-synthesized 1T' few-layered WTe2 NRs is further proposed, which may inspire the growth strategies for other 1D nanostructures. Our findings suggest that ML is a powerful and efficient approach to aid the synthesis of 1D nanostructures, opening up new opportunities for intelligent material development.

cond-mat.mtrl-sci