SearcharxivSearch

arXiv subjects

Zuyuan He

Publications and source records attributed to Zuyuan He.

15 recordsLinked to original sources

A cross-modal pre-training framework with video data for improving performance and generalization of distributed acoustic sensing

Fiber-optic distributed acoustic sensing (DAS) has emerged as a critical Internet-of-Things (IoT) sensing technology with broad industrial applications. However, the two-dimensional spatial-temporal morphology of DAS signals presents analytical challenges where conventional methods prove suboptimal, while being well-suited for deep learning approaches. Although our previous work, DAS Masked Autoencoder (DAS-MAE), established state-of-the-art performance and generalization without labels, it is not satisfactory in frequency analysis in temporal-dominated DAS data. Moreover, the limitation of effective training data fails to address the substantial data requirements inherent to Transformer architectures in DAS-MAE. To overcome these limitations, we present an enhanced framework incorporating short-time Fourier transform (STFT) for explicit temporal-frequency feature extraction and pioneering video-to-DAS cross-modal pre-training to mitigate data constraints. This approach learns high-level representations (e.g., event classification) through label-free reconstruction tasks. Experimental results demonstrate transformative improvements: 0.1% error rate in few-shot classification (90.9% relative improvement over DAS-MAE) and 4.7% recognition error in external damage prevention applications (75.4% improvement over from-scratch training). As the first work to pioneer video-to-DAS cross-modal pre-training, available training resources are expanded by bridging computer vision and distributed sensing areas. The enhanced performance and generalization facilitate DAS deployment across diverse industrial scenarios while advancing cross-modal representation learning for industrial IoT sensing.

eess.SP

Nanosecond-latency all-optical fiber sensing with in-sensor computing

Optical fiber sensing plays a crucial role in modern measurement systems and holds significant promise for a wide range of applications. This potential, though, has been fundamentally constrained by the intrinsic latency and power limitations associated with electronic signal processing. Here, we propose an all-optical fiber sensing architecture with in-sensor computing (AOFS-IC) that achieves fully optical-domain sensing signal demodulation at the speed of light. By integrating a scattering medium with an optimized diffractive optical network, AOFS-IC enables linear mapping of physical perturbations to detected intensity, and sensing results can be directly read out without electronic processing. The proposed system maintains high accuracy across various sensing tasks, providing sub-nano strain resolution and 100% torsional angle classification accuracy, as well as multiplexed sensing of multiple physical quantities, and performing multi-degree-of-freedom robot arm monitoring. AOFS-IC eliminates computing hardware requirements while providing <3 ns demodulation delay, which is more than 2 orders of magnitude faster than conventional fiber optic sensing systems. This work demonstrates the potential of next-generation optical sensing systems empowered by all-optical computing, and paves the way for expanded applications of fiber sensing through the integration of fully optical components, ultrafast measurement speed, and low power consumption.

physics.optics

Resonant microtaper leaky-mode computational spectropolarimetry with tens of femtometers spectral resolution and full stokes measurement

Emerging computational measurement techniques for acquiring multi-dimensional optical field information, such as spectrum and polarization, are rapidly advancing and offer promising solutions for realizing high-performance miniature systems. The performance of these computational measurement approaches is critically influenced by the choice of random media, yet a general framework for evaluating different implementations remains absent. Here, we propose a universal analytical model for computational measurement systems and reveal that the system resolution is fundamentally determined by the maximum optical path difference (OPD) permitted within the random medium. Building on this theoretical foundation, we present a resonant leaky-mode (RLM) spectropolarimeter that achieves a record high resolution-footprint-product metric. The RLM spectropolarimeter leverages the complex coupling between leaky modes in a tapered coreless optical fiber and whispering-gallery modes (WGM) of microsphere to significantly enhance the maximum OPD within a compact footprint. We simultaneously achieve an ultrahigh spectral resolution of 0.02 pm, a spectral measurement bandwidth of 150 nm, and full-Stokes polarization measurement with an accuracy of $4.732 \times 10^{-6}$, all within a sub-square-millimeter footprint. The proposed theoretical model clarifies the key factors governing the performance of computational measurement systems based on random media and may inspires novel design of advanced computational measurement systems for optical field. The demonstrated RLM spectropolarimeter offers a potential approach for highly integrated, high-performance multi-dimensional optical field measurement.

physics.optics

Physics-informed network paradigm with data generation and background noise removal for diverse distributed acoustic sensing applications

Distributed acoustic sensing (DAS) has attracted considerable attention across various fields and artificial intelligence (AI) technology plays an important role in DAS applications to realize event recognition and denoising. Existing AI models require real-world data (RWD), whether labeled or not, for training, which is contradictory to the fact of limited available event data in real-world scenarios. Here, a physics-informed DAS neural network paradigm is proposed, which does not need real-world events data for training. By physically modeling target events and the constraints of real world and DAS system, physical functions are derived to train a generative network for generation of DAS events data. DAS debackground net is trained by using the generated DAS events data to eliminate background noise in DAS data. The effectiveness of the proposed paradigm is verified in event identification application based on a public dataset of DAS spatiotemporal data and in belt conveyor fault monitoring application based on DAS time-frequency data, and achieved comparable or better performance than data-driven networks trained with RWD. Owing to the introduction of physical information and capability of background noise removal, the paradigm demonstrates generalization in same application on different sites. A fault diagnosis accuracy of 91.8% is achieved in belt conveyor field with networks which transferred from simulation test site without any fault events data of test site and field for training. The proposed paradigm is a prospective solution to address significant obstacles of data acquisition and intense noise in practical DAS applications and explore more potential fields for DAS.

cs.LG

DAS-MAE: A self-supervised pre-training framework for universal and high-performance representation learning of distributed fiber-optic acoustic sensing

Distributed fiber-optic acoustic sensing (DAS) has emerged as a transformative approach for distributed vibration measurement with high spatial resolution and long measurement range while maintaining cost-efficiency. However, the two-dimensional spatial-temporal DAS signals present analytical challenges. The abstract signal morphology lacking intuitive physical correspondence complicates human interpretation, and its unique spatial-temporal coupling renders conventional image processing methods suboptimal. This study investigates spatial-temporal characteristics and proposes a self-supervised pre-training framework that learns signals' representations through a mask-reconstruction task. This framework is named the DAS Masked AutoEncoder (DAS-MAE). The DAS-MAE learns high-level representations (e.g., event class) without using labels. It achieves up to 1% error and 64.5% relative improvement (RI) over the semi-supervised baseline in few-shot classification tasks. In a practical external damage prevention application, DAS-MAE attains a 5.0% recognition error, marking a 75.7% RI over supervised training from scratch. These results demonstrate the high-performance and universal representations learned by the DAS-MAE framework, highlighting its potential as a foundation model for analyzing massive unlabeled DAS signals.

eess.SP

Waveguide Superlattices with Artificial Gauge Field Towards Colorless and Crosstalkless Ultrahigh-Density Photonic Integration

Dense waveguides are the basic building blocks for photonic integrated circuits (PIC). Due to the rapidly increasing scale of PIC chips, high-density integration of waveguide arrays working with low crosstalk over broadband wavelength range is highly desired. However, the sub-wavelength regime of such structures has not been adequately explored in practice. Herein, we proposed a waveguide superlattice design leveraging the artificial gauge field (AGF) mechanism, corresponding to the quantum analog of field-induced n-photon resonances in semiconductor superlattices. This approach experimentally achieves -24 dB crosstalk suppression with an ultra-broad transmission bandwidth over 500 nm for dual polarizations. The fabricated waveguide superlattices support high-speed signal transmission of 112 Gbit/s with high-fidelity signal-to-noise ratio profiles and bit error rates. This design, featuring a silica upper cladding, is compatible with standard metal back end-of-the-line (BEOL) processes. Based on such a fundamental structure that can be readily transferred to other platforms, passive and active devices over versatile platforms can be realized with a significantly shrunk on-chip footprint, thus it holds great promise for significant reduction of the power consumption and cost in PICs.

physics.optics

Hierarchical Generative Network for Face Morphing Attacks

Face morphing attacks circumvent face recognition systems (FRSs) by creating a morphed image that contains multiple identities. However, existing face morphing attack methods either sacrifice image quality or compromise the identity preservation capability. Consequently, these attacks fail to bypass FRSs verification well while still managing to deceive human observers. These methods typically rely on global information from contributing images, ignoring the detailed information from effective facial regions. To address the above issues, we propose a novel morphing attack method to improve the quality of morphed images and better preserve the contributing identities. Our proposed method leverages the hierarchical generative network to capture both local detailed and global consistency information. Additionally, a mask-guided image blending module is dedicated to removing artifacts from areas outside the face to improve the image's visual quality. The proposed attack method is compared to state-of-the-art methods on three public datasets in terms of FRSs' vulnerability, attack detectability, and image quality. The results show our method's potential threat of deceiving FRSs while being capable of passing multiple morphing attack detection (MAD) scenarios.

cs.CV

Optimal-Landmark-Guided Image Blending for Face Morphing Attacks

In this paper, we propose a novel approach for conducting face morphing attacks, which utilizes optimal-landmark-guided image blending. Current face morphing attacks can be categorized into landmark-based and generation-based approaches. Landmark-based methods use geometric transformations to warp facial regions according to averaged landmarks but often produce morphed images with poor visual quality. Generation-based methods, which employ generation models to blend multiple face images, can achieve better visual quality but are often unsuccessful in generating morphed images that can effectively evade state-of-the-art face recognition systems~(FRSs). Our proposed method overcomes the limitations of previous approaches by optimizing the morphing landmarks and using Graph Convolutional Networks (GCNs) to combine landmark and appearance features. We model facial landmarks as nodes in a bipartite graph that is fully connected and utilize GCNs to simulate their spatial and structural relationships. The aim is to capture variations in facial shape and enable accurate manipulation of facial appearance features during the warping process, resulting in morphed facial images that are highly realistic and visually faithful. Experiments on two public datasets prove that our method inherits the advantages of previous landmark-based and generation-based methods and generates morphed images with higher quality, posing a more significant threat to state-of-the-art FRSs.

cs.CV

20736-node Weighted Max-Cut Problem Solving by Quadrature Photonic Spatial Ising Machine

To tackle challenging combinatorial optimization problems, analog computing machines based on the nature-inspired Ising model are attracting increasing attentions in order to disruptively overcome the impending limitations on conventional electronic computers. Photonic spatial Ising machine has become an unique and primitive solution with all-to-all connections to solve large-scale Max-cut problems. However, spin configuration and flipping requires two independent sets of spatial light modulators (SLMs) for amplitude and phase modulation, which will lead to tremendous engineering difficulty of optical alignment and coupling. We report a novel quadrature photonic spatial-Euler Ising machine to realize large-scale and flexible spin-interaction configuration and spin-flip in a single spatial light modulator, and develop a noise enhancement approach by adding digital white noise onto detected optical signals. We experimentally show that such proposal accelerates solving (un)weighted, (non)fully connected, 20736-node Max-cut problems, which offers obvious advantages over simulation and heuristic algorithm results in digital computers.

cs.ET

General Spatial Photonic Ising Machine Based on Interaction Matrix Eigendecomposition Method

The spatial photonic Ising machine has achieved remarkable advancements in solving combinatorial optimization problems. However, it still remains a huge challenge to flexibly mapping an arbitrary problem to Ising model. In this paper, we propose a general spatial photonic Ising machine based on interaction matrix eigendecomposition method. Arbitrary interaction matrix can be configured in the two-dimensional Fourier transformation based spatial photonic Ising model by using values generated by matrix eigendecomposition. The error in the structural representation of the Hamiltonian decreases substantially with the growing number of eigenvalues utilized to form the Ising machine. In combination with the optimization algorithm, as low as 65% of the eigenvalues is required by intensity modulation to guarantee the best probability of optimal solution for a 20-vertex graph Max-cut problem, and this probability decreases to below 20% for zero best chance. Our work provides a viable approach for spatial photonic Ising machines to solve arbitrary combinatorial optimization problems with the help of multi-dimensional optical property.

cs.ET

Whispering-gallery-mode barcode-based broadband sub-femtometer-resolution spectroscopy with an electro-optic frequency comb

Spectroscopy is the basic tool for studying molecular physics and realizing bio-chemical sensing. However, it is challenging to realize sub-femtometer resolution spectroscopy over broad bandwidth. In this paper, broadband and high-resolution spectroscopy with calibrated optical frequency is demonstrated by bridging the fields of speckle patterns and electro-optic frequency comb (EOFC). A novel wavemeter based on whispering-gallery-mode (WGM) speckles (or WGM barcodes) is proposed to link the frequency of a tunable continuous-wave (CW) laser to an optical reference provided by an ultra-stable laser. The ultra-fine comb lines generated from the CW laser sample the spectrum with sub-femtometer resolution. Measurement bandwidth is far extended by performing sequential acquisitions, since the centre optical frequency of EOFC is absolutely determined by WGM speckle-based wavemter. This approach fully utilizes the advantages of two fields to realize 0.8-fm resolution with a fiber laser and 80-nm bandwidth with an external cavity diode laser. The spectroscopic measurements of an ultrahigh-Q cavity and the HCN gas absorption is demonstrated, which shows the potentials of this compact system with high resolution and broad bandwidth for more applications.

physics.optics

Quadrature Photonic Spatial Ising Machine

The mining in physics and biology for accelerating the hardcore algorithm to solve non-deterministic polynomial (NP) hard problems has inspired a great amount of special-purpose ma-chine models. Ising machine has become an efficient solver for various combinatorial optimizationproblems. As a computing accelerator, large-scale photonic spatial Ising machine have great advan-tages and potentials due to excellent scalability and compact system. However, current fundamentallimitation of photonic spatial Ising machine is the configuration flexibility of problem implementationin the accelerator model. Arbitrary spin interactions is highly desired for solving various NP hardproblems. Moreover, the absence of external magnetic field in the proposed photonic Ising machinewill further narrow the freedom to map the optimization applications. In this paper, we propose anovel quadrature photonic spatial Ising machine to break through the limitation of photonic Isingaccelerator by synchronous phase manipulation in two and three sections. Max-cut problem solutionwith graph order of 100 and density from 0.5 to 1 is experimentally demonstrated after almost 100iterations. We derive and verify using simulation the solution for Max-cut problem with more than1600 nodes and the system tolerance for light misalignment. Moreover, vertex cover problem, modeled as an Ising model with external magnetic field, has been successfully implemented to achievethe optimal solution. Our work suggests flexible problem solution by large-scale photonic spatialIsing machine.

cs.ET

High-speed silicon microring modulator at 2-um waveband

We demonstrated a silicon integrated microring modulator working at the 2-um waveband with an L-shaped PN junction. 15-GHz 3-dB electro-optic bandwidth and <1 Vcm modulation efficiency for 45-Gbps NRZ-OOK signaling is achieved at 1960 nm.

physics.app-ph

Photonic Convolution Neural Network Based on Interleaved Time-Wavelength Modulation

Convolution neural network (CNN), as one of the most powerful and popular technologies, has achieved remarkable progress for image and video classification since its invention in 1989. However, with the high definition video-data explosion, convolution layers in the CNN architecture will occupy a great amount of computing time and memory resources due to high computation complexity of matrix multiply accumulate operation. In this paper, a novel integrated photonic CNN is proposed based on double correlation operations through interleaved time-wavelength modulation. Micro-ring based multi-wavelength manipulation and single dispersion medium are utilized to realize convolution operation and replace the conventional optical delay lines. 200 images are tested in MNIST datasets with accuracy of 85.5% in our photonic CNN versus 86.5% in 64-bit computer.We also analyze the computing error of photonic CNN caused by various micro-ring parameters, operation baud rates and the characteristics of micro-ring weighting bank. Furthermore, a tensor processing unit based on 4x4 mesh with 1.2 TOPS (operation per second when 100% utilization) computing capability at 20G baud rate is proposed and analyzed to form a paralleled photonic CNN.

cs.ET

Arbitrarily routed mode-division multiplexed photonic circuits for dense integration

Mode-division multiplexing (MDM) is becoming an enabling technique for large-capacity data communications via encoding the information on orthogonal guiding modes. However, the on-chip routing of a multimode waveguide occupies too large chip area due to the constraints on inter-mode cross talk and mode leakage. Very recently, many efforts have been made to shrink the footprint of individual element like bending and crossing, but the devices still occupy >10x10 um2 footprint for three-mode multiplexed signals and the high-speed signal transmission has not been demonstrated yet. In this work, we demonstrate the first MDM circuits based on digitized meta-structures which have extremely compact footprints. The radius for a three-mode bending is only 3.9 μm and the footprint of a crossing is only 8x8um2. The 3x100 Gbit/s mode-multiplexed signals are arbitrarily routed through the circuits consists of many sharp bends and compact crossing with a bit error rate under forward error correction limit. This work is a significant step towards the large-scale and dense integration of MDM photonic integrated circuits.

physics.app-ph