A General Implementation Details For our Atari games and DeepMind Control Suite experiments, we largely follow DrQ [ 33

Neural Information Processing Systems 

The replay buffer size is 100K. This procedure is repeated every time an image is sampled from the replay buffer. Batch size used in both RL and representation learning is 512. The corresponding hyperparameters used in Atari experiments are shown in Table 7 and Table 8. The action repeat hyperparameters are show in Table 6.

Similar Docs  Excel Report  more

TitleSimilaritySource
None found