Searcharxiv⌕ Search

arXiv subjects

Seung Joon Lee

Publications and source records attributed to Seung Joon Lee.

3 recordsLinked to original sources

ALIGN: Advanced Query Initialization with LiDAR-Image Guidance for Occlusion-Robust 3D Object Detection

Recent query-based 3D object detection methods using camera and LiDAR inputs have shown strong performance, but existing query initialization strategies,such as random sampling or BEV heatmap-based sampling, often result in inefficient query usage and reduced accuracy, particularly for occluded or crowded objects. To address this limitation, we propose ALIGN (Advanced query initialization with LiDAR and Image GuidaNce), a novel approach for occlusion-robust, object-aware query initialization. Our model consists of three key components: (i) Occlusion-aware Center Estimation (OCE), which integrates LiDAR geometry and image semantics to estimate object centers accurately (ii) Adaptive Neighbor Sampling (ANS), which generates object candidates from LiDAR clustering and supplements each object by sampling spatially and semantically aligned points around it and (iii) Dynamic Query Balancing (DQB), which adaptively balances queries between foreground and background regions. Our extensive experiments on the nuScenes benchmark demonstrate that ALIGN consistently improves performance across multiple state-of-the-art detectors, achieving gains of up to +0.9 mAP and +1.2 NDS, particularly in challenging scenes with occlusions or dense crowds. Our code will be publicly available upon publication.

cs.CV↗

Image-Guided Semantic Pseudo-LiDAR Point Generation for 3D Object Detection

In autonomous driving scenarios, accurate perception is becoming an even more critical task for safe navigation. While LiDAR provides precise spatial data, its inherent sparsity makes it difficult to detect small or distant objects. Existing methods try to address this by generating additional points within a Region of Interest (RoI), but relying on LiDAR alone often leads to false positives and a failure to recover meaningful structures. To address these limitations, we propose Image-Guided Semantic Pseudo-LiDAR Point Generation model, called ImagePG, a novel framework that leverages rich RGB image features to generate dense and semantically meaningful 3D points. Our framework includes an Image-Guided RoI Points Generation (IG-RPG) module, which creates pseudo-points guided by image features, and an Image-Aware Occupancy Prediction Network (I-OPN), which provides spatial priors to guide point placement. A multi-stage refinement (MR) module further enhances point quality and detection robustness. To the best of our knowledge, ImagePG is the first method to directly leverage image features for point generation. Extensive experiments on the KITTI and Waymo datasets demonstrate that ImagePG significantly improves the detection of small and distant objects like pedestrians and cyclists, reducing false positives by nearly 50%. On the KITTI benchmark, our framework improves mAP by +1.38%p (car), +7.91%p (pedestrian), and +5.21%p (cyclist) on the test set over the baseline, achieving state-of-the-art cyclist performance on the KITTI leaderboard. The code is available at: https://github.com/MS-LIMA/ImagePG

cs.CV↗

Development of Thin-Gap GEM-μRWELL Hybrid Detectors

Micro Pattern Gaseous Detectors (MPGDs) are used for tracking in High Energy Physics and Nuclear Physics because of their large area, excellent spatial resolution capabilities and low cost. However, for high energy charged particles impacting at a large angle with respect to the axis perpendicular to detector plane, the spatial resolution degrades significantly because of the long trail of ionization charges produced in clusters all along the track in the drift region of the detector. The long ionization charge trail results in registering hits from large number of strips in the readout plane which makes it challenging to precisely reconstruct the particle position using simple center of gravity algorithm. As a result, the larger the drift gap, the more severe the deterioration of spatial resolution for inclined tracks. For the same reason, the position resolution is also severely degraded in a large magnetic field, where the Lorentz E {\times} B effect causes the ionization charges to follow a curved and longer path in the detector gas volume. In this paper, we report on the development of thin-gap MPGDs as a way to maintain excellent spatial resolution capabilities of MPGD detectors over a wide angular range of incoming particles. In a thin-gap MPGD, the thickness of the gas volume in the drift region is reduced from typically {\sim} 3 mm to {\sim} 1 mm or less. We present preliminary test beam results demonstrating the improvement in spatial resolution from {\sim} 400 μm with a standard 3 mm gap μRWELL prototype to {\sim} 140 μm with a double amplification GEM-μRWELL thin-gap hybrid detector. We also discuss the impact of a thin-gap drift volume on other aspects of the performance of MPGD technologies such as the efficiency and detector stability.

physics.ins-det↗