in environments that are only sparsely rewarding, curiosity is a very big deal. also, lazy robots who decide to watch TV instead of exploration :D
https://blog.openai.com/reinforcement-learning-with-prediction-based-rewards/
Post #563
228
LI Linkstream @linkstream · 169 subscribers