postponed for a week due to the wind!
We come together to discuss advances in computer science, ground them down to applications, and see where to the wind is blowing. Read the papers at your own time and come discuss them with us.
The plan:
- How to compare RL algorithms:
- The Primacy Bias in RL: when resets work, and when they don't
- Resets in PPO: what we tried (spoiler: nothing worked)
- Heavy Priming experiment: overfitting on one batch at the beginning of the training. Differences between SAC and PPO, the probable cause of PPO stability.
⏱ 18:00-20:00 Thursday, 28 December
📍 F0RTHSP4CE,
🗣
👮
