SearcharxivSearch

arXiv subjects

Haodi Wu

Publications and source records attributed to Haodi Wu.

2 recordsLinked to original sources

SDTrack: A Baseline for Event-based Tracking via Spiking Neural Networks

Event cameras provide superior temporal resolution, dynamic range, energy efficiency, and pixel bandwidth. Spiking Neural Networks (SNNs) naturally complement event data through discrete spike signals, making them ideal for event-based tracking. However, current approaches combining Artificial Neural Networks (ANNs) and SNNs suffer from suboptimal architectures that compromise energy efficiency and limit tracking performance. To address these limitations, we propose the first Transformer-based \textbf{S}pike-\textbf{D}riven \textbf{T}racking (SDTrack) pipeline. It incorporates a novel event frame aggregation method called Global Trajectory Prompt (GTP) and a Transformer-based tracker. The GTP method effectively captures global trajectory information and aggregates it with event streams into event frames to enhance spatiotemporal representation. The Transformer-based tracker comprises a fully spike-driven SNN backbone and a simple tracking head. The SDTrack pipeline operates end-to-end without data augmentation or post-processing. Extensive experiments demonstrate that our SDTrack-Tiny pipeline achieves competitive accuracy with only 19.61$M$ parameters and 8.16$mJ$ energy consumption, while our Base version achieves state-of-the-art accuracy across three datasets. Our work establishes a solid foundation for future neuromorphic vision research.

cs.NE

DIRECT-Net: a unified mutual-domain material decomposition network for quantitative dual-energy CT imaging

By acquiring two sets of tomographic measurements at distinct X-ray spectra, the dual-energy CT (DECT) enables quantitative material-specific imaging. However, the conventionally decomposed material basis images may encounter severe image noise amplification and artifacts, resulting in degraded image quality and decreased quantitative accuracy. Iterative DECT image reconstruction algorithms incorporating either the sinogram or the CT image prior information have shown potential advantages in noise and artifact suppression, but with the expense of large computational resource, prolonged reconstruction time, and tedious manual selections of algorithm parameters. To partially overcome these limitations, we develop a domain-transformation enabled end-to-end deep convolutional neural network (DIRECT-Net) to perform high quality DECT material decomposition. Specifically, the proposed DIRECT-Net has immediate accesses to mutual-domain data, and utilizes stacked convolution neural network (CNN) layers for noise reduction and material decomposition. The training data are numerically simulated based on the underlying physics of DECT imaging.The XCAT digital phantom, iodine solutions phantom, and biological specimen are used to validate the performance of DIRECT-Net. The qualitative and quantitative results demonstrate that this newly developed DIRECT-Net is promising in suppressing noise, improving image accuracy, and reducing computation time for future DECT imaging.

physics.med-ph