3 ms·
That’s not how training pipelines work, and would be extremely wasteful for the biggest cost center as well.
by stingraycharles 2mo ago
That’s not how training pipelines work, and would be extremely wasteful for the biggest cost center as well.
- jluysvi 2mo agoHow about expanding a little instead of just saying "that's not how training pipelines work".
- vorticalbox 2mo agopretty sure at this point no one is retraining from zero they have their big model and they fine tune it. different training makes a different version (agent, info sec etc).
- jychang 2mo agoThe GPT-5.5 and 5.6 Spud pretrain is a fresh pretrain run. OpenAI has the Doug/Astro pretrain coming up next.
- ac29 2mo agoI was under the impression labs released post-training checkpoints fairly often? So Model N/N+1 might literally have had the exact the same pretraining run and only differ on how much/what kind of postraining they got