SearcharxivSearch

arXiv subjects

Qidong Wang

Publications and source records attributed to Qidong Wang.

7 recordsLinked to original sources

Large-Aperture All-Solid-State Cascaded Liquid-Crystal Beam Steering for High-Resolution Wide-Field Imaging

High-resolution wide-field imaging is essential for applications requiring simultaneous global coverage and local detail, yet conventional approaches face a fundamental trade-off: wide-FOV cameras sacrifice spatial sampling density by distributing finite detector pixels over a broad angular range, while telephoto systems resolve fine features at the cost of scene coverage. Beam-steering devices can mitigate this trade-off but are currently limited in achieving simultaneously all-solid-state, large aperture, and high-speed operation. Here, we report an all-solid-state large-aperture cascaded liquid-crystal beam-steering (CaLiBS) imaging system that extends the effective angular range of a high-resolution narrow-FOV camera by electrically steering sub-FOVs. The CaLiBS module comprises cascaded liquid crystal waveplates and liquid crystal Pancharatnam-Berry phase gratings; a theoretical voltage-prediction model with a hierarchical search algorithm enables efficient calibration under oblique incidence and 10 times faster calibration speed compared with conventional methods. The calibrated system addresses sub-FOVs across 30.3° * 30.3° at 2° intervals with diffraction efficiency above 60%. Sequential sub-FOV acquisition reconstructs a 34.7 * 34.7 composite image, an 8.6-fold enhancement in spatial-bandwidth product over a single-shot wide-FOV camera using the same detector. Combined with object tracking methods, sub-FOV switching further enables high-resolution tracking of moving vehicles within the wide-area scene. This cascaded LC architecture offers a scalable pathway toward compact, vibration-free, and high-resolution wide-field observation.

physics.optics

V-SEAM: Visual Semantic Editing and Attention Modulating for Causal Interpretability of Vision-Language Models

Recent advances in causal interpretability have extended from language models to vision-language models (VLMs), seeking to reveal their internal mechanisms through input interventions. While textual interventions often target semantics, visual interventions typically rely on coarse pixel-level perturbations, limiting semantic insights on multimodal integration. In this study, we introduce V-SEAM, a novel framework that combines Visual Semantic Editing and Attention Modulating for causal interpretation of VLMs. V-SEAM enables concept-level visual manipulations and identifies attention heads with positive or negative contributions to predictions across three semantic levels: objects, attributes, and relationships. We observe that positive heads are often shared within the same semantic level but vary across levels, while negative heads tend to generalize broadly. Finally, we introduce an automatic method to modulate key head embeddings, demonstrating enhanced performance for both LLaVA and InstructBLIP across three diverse VQA benchmarks. Our data and code are released at: https://github.com/petergit1/V-SEAM.

cs.CL

From Heads to Neurons: Causal Attribution and Steering in Multi-Task Vision-Language Models

Recent work has increasingly explored neuron-level interpretation in vision-language models (VLMs) to identify neurons critical to final predictions. However, existing neuron analyses generally focus on single tasks, limiting the comparability of neuron importance across tasks. Moreover, ranking strategies tend to score neurons in isolation, overlooking how task-dependent information pathways shape the write-in effects of feed-forward network (FFN) neurons. This oversight can exacerbate neuron polysemanticity in multi-task settings, introducing noise into the identification and intervention of task-critical neurons. In this study, we propose HONES (Head-Oriented Neuron Explanation & Steering), a gradient-free framework for task-aware neuron attribution and steering in multi-task VLMs. HONES ranks FFN neurons by their causal write-in contributions conditioned on task-relevant attention heads, and further modulates salient neurons via lightweight scaling. Experiments on four diverse multimodal tasks and two popular VLMs show that HONES outperforms existing methods in identifying task-critical neurons and improves model performance after steering. Our source code is released at: https://github.com/petergit1/HONES.

cs.CV

Monte Carlo Simulation of Angular Response of GRID Detectors for GRID Mission

The Gamma-Ray Integrated Detectors (GRID) are a space science mission that employs compact gamma-ray detectors mounted on NanoSats in low Earth orbit (LEO) to monitor the transient gamma-ray sky. Owing to the unpredictability of the time and location of gamma-ray bursts (GRBs), obtaining the photon responses of gamma-ray detectors at various incident angles is important for the scientific analysis of GRB data captured by GRID detectors. For this purpose, a dedicated Monte Carlo simulation framework has been developed for GRID detectors. By simulating each GRID detector and the NanoSat carrying it, the spectral energy response, detection efficiency, and other angular responses of each detector for photons with different incident angles and energies can be obtained within this framework. The accuracy of these simulations has been corroborated through on-ground calibration, and the derived angular responses have been successfully applied to the data analysis of recorded GRBs.

astro-ph.IM

Development of silicon interposer: towards an ultralow radioactivity background photodetector system

It is of great importance to develop a photodetector system with an ultralow radioactivity background in rare event searches. Silicon photomultipliers (SiPMs) and application-specific integrated circuits (ASICs) are two ideal candidates for low background photosensors and readout electronics, respectively, because they are mainly composed of silicon, which can achieve good radio-purity without considerable extra effort. However, interposers, used to provide mechanical support and signal routes between the photosensor and the electronics, are a bottleneck in building ultralow background photodetectors. Silicon and quartz are two candidates to construct the low background interposer because of their good radio-purity; nevertheless, it is non-trivial to produce through silicon vias (TSV) or through quartz vias (TQV) on the large area silicon or quartz wafer. In this work, based on double-sided TSV interconnect technology, we developed the first prototype of a silicon interposer with a size of 10~cm$\times$10~cm and a thickness of 320~$μ$m. The electrical properties of the interposer are carefully evaluated at room temperature, and its performance is also examined at -110~$^\circ$C with an integrated SiPM on the interposer. The testing results reveal quite promising performance of the prototype, and the single photoelectron signals can be clearly observed from the SiPM. The features of the observed signals are comparable with those from the SiPM mounted on a normal FR4-based PCB. Based on the success of the silicon interposer prototype, we started the follow-up studies that aimed to further improve the performance and yield of the silicon interposer, and eventually to provide a solution for building an ultralow background photodetector system.

physics.ins-det

Search for non-Newtonian interactions at micrometer scale with a levitated test mass

We report on a search for non-Newtonian forces that couple to mass, with a characteristic scale of ${\sim}10~μ$m, using an optically levitated microsphere as a precision force sensor. A silica microsphere trapped in an upward-propagating, single-beam, optical tweezer is utilized to probe for interactions sourced from a nanofabricated attractor mass with a density modulation brought into close proximity to the microsphere and driven along the axis of periodic density in order to excite an oscillating response. We obtain force sensitivity of ${\lesssim}10^{-16}~\rm{N}/\sqrt{\rm{Hz}}$. Separately searching for attractive and repulsive forces results in the constraint on a new Yukawa interaction of $|α| \gtrsim 10^8$ for $λ> 10~μ$m. This is the first test of the inverse-square law using an optically levitated test mass of dimensions comparable to $λ$, a complementary method subject to a different set of systematic effects compared to more established techniques.

hep-ex

Three-dimensional force-field microscopy with optically levitated microspheres

We report on the use of 4.7-$μ$m-diameter, optically levitated, charged microspheres to image the three-dimensional force field produced by charge distributions on an Au-coated, microfabricated Si beam in vacuum. An upward-propagating, single-beam optical trap, combined with an interferometric imaging technique, provides optimal access to the microspheres for microscopy. In this demonstration, the Au-coated surface of the Si beam can be brought as close as ${\sim}10~μ$m from the center of the microsphere while forces are simultaneously measured along all three orthogonal axes, fully mapping the vector force field over a total volume of ${\sim}10^6~μ$m$^3$. We report a force sensitivity of $(2.5 \pm 1.0) \times 10^{-17}~{\rm N / \sqrt{Hz}}$, in each of the three degrees of freedom, with a linear response to up to ${\sim}10^{-13}~{\rm N}$. While we discuss the case of mapping static electric fields using charged microspheres, it is expected that the technique can be extended to other force fields, using microspheres with different properties.

physics.ins-det