TY - RPRT TI - $QD$-Learning: A Collaborative Distributed Strategy for Multi-Agent Reinforcement Learning Through Consensus + Innovations AU - Soummya Kar AU - Jose' M. F. Moura AU - H. Vincent Poor PY - 2012 DO - 10.1109/tsp.2013.2241057 UR - https://arxiv.org/abs/1205.0047 ID - 1205.0047 ER -