TY - RPRT TI - An Adaptive Method for Contextual Stochastic Multi-armed Bandits with Rewards Generated by a Linear Dynamical System AU - Jonathan Gornet AU - Mehdi Hosseinzadeh AU - Bruno Sinopoli PY - 2025 UR - https://arxiv.org/abs/2406.10418 ID - 2406.10418 ER -