Searcharxiv⌕ Search

arXiv · 2610.06565

Blind Beamforming for Intelligent Reflecting Surfaces: A Gradient Bandit Approach

Abstract

The beamforming problem of intelligent reflecting surface (IRS) has been extensively considered from an optimization perspective assuming that channel state information (CSI) is available. However, the reality is that the existing prototypes seldom follow this model-based approach because channel estimation is technically difficult and costly for the network protocols and hardware to date. A recent trend is to perform beamforming blindly without channel knowledge. This work looks at blind beamforming from a reinforcement learning point of view. We first show that the existing blind beamforming method boils down to a special case of the greedy algorithm in the reinforcement learning context. We analyze the resulting cumulative regret, and further propose an upper approximation to facilitate the optimization of the exploration probability. Moreover, we show that a gradient sampling scheme can improve the efficiency of reinforcement learning as compared to the uniform sampling scheme adopted in the existing blind beamforming method. We further verify the convergence of the proposed gradient sampling scheme by showing that it is a stochastic approximation to gradient ascent. Finally, we physically implement the proposed method in a real-world prototype system. Our field test results show that, as compared to the existing blind beamforming method, the proposed gradient sampling boosts the average signal-to-noise ratio (SNR) by more than 5 dB with 1000 samples.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Wenhai Lai, Mingxiao Li, Kaiming Shen. 2026-10-05. Blind Beamforming for Intelligent Reflecting Surfaces: A Gradient Bandit Approach. https://arxiv.org/abs/2610.06565

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Resource-Efficient Wi-Fi CSI-Based Sensing via Exploiting the Age of Samples

Wi-Fi channel state information (CSI)-based sensing must coexist with data communications, limiting the availability of temporally-dense CSI measurements. We formulate CSI-based human activity and identity recognition under an average sensing budget that limits the fraction of CSI measurement and reporting opportunities within a sensing session. The budget captures sensing-communication resource sharing, packet loss, and traffic-induced irregularity, which we model using deterministic (accumulated) and stochastic (Bernoulli) sampling policies. We propose a low-cost, age-aware WiFi sensing framework that encodes the age of each retained CSI sample and multiplicatively fuses it with the CSI embedding. On the NTU-Fi human activity recognition and person identification datasets, the proposed model outperforms both a CSI-only baseline and the time-aware attention model of the UniFi benchmark across most operating regimes. For person identification, it improves over UniFi by more than 10 percentage points, with the largest gains under strict sensing budgets.

eess.SP↗

Graph Learning for Cross-Subject, Cross-Population EEG Emotion Decoding and Model-Derived Spatial-Spectral Neural Signatures

Cross subject emotion decoding from electroencephalography EEG requires representations that accommodate individual variability while preserving spatial spectral structure for interpretation. This study introduces EmoDiPyraTrans, a differential graph Transformer that integrates adaptive graph recurrence, differential attention, pyramid fusion and distribution regularization over sequential relative power spectral density graphs. Across SEED, FACED, MAHNOB HCI, DEAP and DREAMER, the model achieved the highest participant mean accuracy and positive class F1 among the evaluated methods, with accuracy and F1 both reaching 0.928 on SEED. On DEP EEG, positive versus neutral accuracy reached 0.802 within healthy controls and 0.704 within participants with depression, compared with 0.591 under healthy to depression transfer and 0.581 with mixed population development. Complementary SEED analyses identified distributed spatial weighting and an alpha centred spectral preference, while configurations averaging six channels retained near full performance. These findings link generalization assessment with model derived candidate signatures to support interpretable EEG emotion decoding, with code available at https://github.com/hdy6438/EmoDiPyraTrans.

eess.SP↗