multiprocessing에서 gym이 오작동하는 경우
tl;dr
- gym>=0.10.6이면 상관x
- multiprocessing으로 학습을 진행시키는 경우 주의
IMPALA: Scalable Distributed Deep-RL with Importance Weighted Actor-Learner Architectures https://arxiv.org/abs/1802.01561
More …https://arxiv.org/abs/1805.11604
More …Tensorflow에서 cross-entropy를 구현하는 법에 대해 좋은
Stack overflow포스트가 있어서 공유
환경은 Python/ Tensorflow(1.13)/ Keras
More …