Paper2016
Asynchronous Methods for Deep Reinforcement Learning
Mnih et al.
Shows that running many actor-learners asynchronously on CPU cores stabilises deep reinforcement learning as well as a replay buffer does, while training faster and off GPU.
link checked 17 Sept 2026FreeIntermediate