SearcharxivSearch

arXiv subjects

Yongjian Luo

Publications and source records attributed to Yongjian Luo.

3 recordsLinked to original sources

CADSpotting: Robust Panoptic Symbol Spotting on Large-Scale CAD Drawings

We introduce CADSpotting, an effective method for panoptic symbol spotting in large-scale architectural CAD drawings. Existing approaches often struggle with symbol diversity, scale variations, and overlapping elements in CAD designs, and typically rely on additional features (e.g., primitive types or graphical layers) to improve performance. CADSpotting overcomes these challenges by representing primitives through densely sampled points with only coordinate attributes, using a unified 3D point cloud model for robust feature learning. To enable accurate segmentation in large drawings, we further propose a novel Sliding Window Aggregation (SWA) technique that combines weighted voting and Non-Maximum Suppression (NMS). Moreover, we introduce LS-CAD, a new large-scale dataset comprising 45 finely annotated floorplans, each covering approximately 1,000 $m^2$, significantly larger than prior benchmarks. LS-CAD will be publicly released to support future research. Experiments on FloorPlanCAD and LS-CAD demonstrate that CADSpotting significantly outperforms existing methods. We also showcase its practical value by enabling automated parametric 3D interior reconstruction directly from raw CAD inputs.

cs.CV

Empowering Feed-Forward Reconstruction Models with Metric Scale via Satellite Images

Feed-forward 3D reconstruction models have recently shown strong generalization across diverse scenes, yet most of them recover geometry only up to an unknown global scale. This scale ambiguity limits their use in applications that require metric understanding of the environment. Existing metric reconstruction methods commonly rely on large-scale metric annotations or accurate camera calibration, both of which are costly or unreliable in many real-world settings. We propose a satellite-guided framework for resolving scale ambiguity in feed-forward 3D reconstruction. The key idea is to use readily available satellite imagery as a global metric reference. Given a coarse camera pose, our method retrieves a local satellite patch and integrates it with a feed-forward reconstruction backbone through bidirectional cross-view interaction. By enforcing consistency between the reconstructed scene and the satellite reference, the model infers absolute scale, refines scene geometry, and estimates camera pose in a metric coordinate frame. Experiments on KITTI, nuScenes, and Oxford RobotCar show consistent improvements in metric depth estimation, multi-view point-cloud reconstruction, and cross-view camera localization, while preserving strong generalization across datasets and geographic regions.

cs.CV

Tripling energy storage density through order-disorder transition induced polar nanoregions in PbZrO3 thin films by ion implantation

Dielectric capacitors are widely used in pulsed power electronic devices due to their ultrahigh power densities and extremely fast charge/discharge speed. To achieve enhanced energy storage density, both maximum polarization (Pmax) and breakdown strength (Eb) need to be improved simultaneously. However, these two key parameters are inversely correlated. In this study, order-disorder transition induced polar nanoregions (PNRs) have been achieved in PbZrO3 thin films by making use of the low-energy ion implantation, enabling us overcome the trade-off between high polarizability and breakdown strength, which leads to the tripling of the energy storage density from 20.5 J/cm3 to 62.3 J/cm3 as well as the great enhancement of breakdown strength. This approach could be extended to other dielectric oxides to improve the energy storage performance, providing a new pathway for tailoring the oxide functionalities.

cond-mat.mtrl-sci