Searcharxiv⌕ Search

arXiv subjects

Rongjin Zhuang

Publications and source records attributed to Rongjin Zhuang.

2 recordsLinked to original sources

Programmable electro-optic frequency comb empowers integrated parallel convolution processing

Integrated photonic convolution processors make optical neural networks (ONNs) a transformative solution for artificial intelligence applications such as machine vision. To enhance the parallelism, throughput, and energy efficiency of ONNs, wavelength multiplexing is widely applied. However, it often encounters the challenges of low compactness, limited scalability, and high weight reconstruction latency. Here, we proposed and demonstrated an integrated photonic processing unit with a parallel convolution computing speed of 1.62 trillion operations per second (TOPS) and a weight reconstruction speed exceeding 38 GHz. This processing unit simultaneously achieves, for the first time, multi-wavelength generation and weight mapping via a single programmable electro-optic (EO) frequency comb, featuring unprecedented compactness, device-footprint independent scalability, and near-unity optical power conversion efficiency (conversion efficiency from input optical power to output weighted comb lines). To demonstrate the reconfigurability and functionality of this processing unit, we implemented image edge detection and object classification based on EO combs obtained using the particle swarm algorithm and an EO comb neural network training framework, respectively. Our programmable EO comb-based processing framework establishes a new paradigm towards the development of low-latency monolithic photonic processors, promising real-time in-sensor learning for autonomous vehicles, intelligent robotics, and drones.

physics.optics↗

Self-supervised Multiplex Consensus Mamba for General Image Fusion

Image fusion integrates complementary information from different modalities to generate high-quality fused images, thereby enhancing downstream tasks such as object detection and semantic segmentation. Unlike task-specific techniques that primarily focus on consolidating inter-modal information, general image fusion needs to address a wide range of tasks while improving performance without increasing complexity. To achieve this, we propose SMC-Mamba, a Self-supervised Multiplex Consensus Mamba framework for general image fusion. Specifically, the Modality-Agnostic Feature Enhancement (MAFE) module preserves fine details through adaptive gating and enhances global representations via spatial-channel and frequency-rotational scanning. The Multiplex Consensus Cross-modal Mamba (MCCM) module enables dynamic collaboration among experts, reaching a consensus to efficiently integrate complementary information from multiple modalities. The cross-modal scanning within MCCM further strengthens feature interactions across modalities, facilitating seamless integration of critical information from both sources. Additionally, we introduce a Bi-level Self-supervised Contrastive Learning Loss (BSCL), which preserves high-frequency information without increasing computational overhead while simultaneously boosting performance in downstream tasks. Extensive experiments demonstrate that our approach outperforms state-of-the-art (SOTA) image fusion algorithms in tasks such as infrared-visible, medical, multi-focus, and multi-exposure fusion, as well as downstream visual tasks.

cs.CV↗