TGViewer
DoomPosting DoomPosting @doomposting · 7.94K subscribers
Post #267010 617
It really does seem pre-training as we know it is beginning to die. Of course, Ilya spoke on this. The pre-training corpus has been exhausted. Now the primary model gains come from RL-able tasks. Of all the people at the labs I speak with, almost no one mentions pre-training anymore; just post-training/RL and continuous learning (test-time-training).

This is why the models have become so much better (on evals at least) with math/coding while seemingly having stagnated on other more general, less RL-able domains.

This is why all the labs are spending so much money on different RL envs; so they can smooth out the model capability 'spikiness'.

🄳🄾🄾🄼🄿🤖🅂🅃🄸🄽🄶
  • 💯 1
More from @doomposting
  1. Oct 7, 2026My friend, myslozbir, a Polish patriot who did reporting and commentary on trans extremist…
  2. Oct 7, 2026I didn’t report my rapist because she was my babysitter and had a great ass 🄳🄾🄾🄼🄿🤖🅂…
  3. Oct 7, 2026Did he really say that? 🄳🄾🄾🄼🄿🤖🅂🅃🄸🄽🄶
  4. Oct 7, 2026A Brooklyn judge told a defendant the Second Amendment doesn’t exist in her courtroom, the…
  5. Oct 7, 2026🄳🄾🄾🄼🄿🤖🅂🅃🄸🄽🄶
  6. Oct 7, 2026🄳🄾🄾🄼🄿🤖🅂🅃🄸🄽🄶
Threads Profile ViewerView any public Threads profile without an account.Open ThreadLook →Writing with AI? Make it sound human.Metric37 rewrites AI drafts so they read naturally. Free AI detector, 1,500 words free.Try Metric37 →