arXiv · 2609.35459
Compressed LLM Reprogramming for Vision-Aided Beam Prediction in Vehicular Networks
Abstract
Large Language Model (LLM) reprogramming-based beam prediction demonstrates strong data efficiency by adapting pretrained language models for vehicle-to-infrastructure (V2I) beam prediction, yet the resulting model complexity makes such approaches impractical for latency-sensitive deployment. We propose LLMBP-Lite, a compact LLM-reprogrammed framework for beam prediction. It leverages structural redundancy through pruning along three complementary dimensions: Transformer depth, source-prototype vocabulary size, and prompt length. Additionally, knowledge distillation can be optionally employed to maintain the pretrained representational capacity after compression. Experiments on the real-world dataset demonstrate that LLMBP-Lite achieves a $65\times$ inference speedup over the uncompressed LLM-based baseline while maintaining prediction accuracy and consistently outperforming recurrent baselines under limited training data. These results demonstrate that the tradeoff between data efficiency and deployment efficiency can be substantially mitigated in vehicular networks.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Kai Dong, Lei Wang, Changyi Li, Sergiy A. Vorobyov, Stefan Werner. 2026-09-28. Compressed LLM Reprogramming for Vision-Aided Beam Prediction in Vehicular Networks. https://arxiv.org/abs/2609.35459
Cite the original work for its findings. Save a collection to share your selection of sources.