2 ms·
Wow no reinforcement learning needed! Amazing to think, all they would've had to do instead is spend a decade or two building towards a 1 trillion parameter tra
by cypherpunks01 3y ago
Wow no reinforcement learning needed! Amazing to think, all they would've had to do instead is spend a decade or two building towards a 1 trillion parameter transformer model and spend 100m or two training it. Then fine-tune the model using.. y'know nevermind : )