Post #11101
3.3K
Forwarded from Just links
Self-Play Pretraining with Zero Data https://arxiv.org/abs/2609.30063
arXiv.org Self-Play Pretraining with Zero Data Advances in language modeling have been driven by scaling pretraining on ever more data. Yet, the training data is still largely curated on the model's behalf. A more general approach to... - 🥴 26
- 🔥 7
- 👍 3
- 🌭 3