Abandoning Objectives: Evolution Through the Search for Novelty Alone
Playing Atari with Deep Reinforcement Learning
Asynchronous Methods for Deep Reinforcement Learning
Evolution Strategies as a Scalable Alternative to Reinforcement Learning
https://arxiv.org/pdf/1712.06567.pdf#page=15&org=uber