A Multi-Agent Off-Policy Actor-Critic Algorithm for Distributed Reinforcement Learning. Issue 2 (2020)