3 ms·
2 PB? They will not come close to training in on that amount. Maybe years from now.
by 7e 4mo ago
2 PB? They will not come close to training in on that amount. Maybe years from now.
- huflungdung 4mo ago[dead]
- Den_VR 4mo agoCould probably LoRA with that
- sgt 4mo agoThink they will not train on the dull 2TB but use that as the data lake to start and then apply a more targeted approach.
- winddude 4mo agoif you read the article 2pb is available as flash storage in the data pipeline, used to dedupe, clean, normalize, etc, for training from 60pb of raw data.