Post #303 5.71K May 28, 2026, 18:03 UTC Отчёты перед инвесторами дуреют от этой прикормки https://arxiv.org/abs/2309.08632 arXiv.org Pretraining on the Test Set Is All You Need Inspired by recent work demonstrating the promise of smaller Transformer-based language models pretrained on carefully curated data, we supercharge such approaches by investing heavily in curating... 😁 24 🔥 13 🦄 3