3 ms·
Is that right? Why the shit would that happen?
by pelagicAustral 16d ago
Is that right? Why the shit would that happen?
- transdev12 16d agoBecause they’ve essentially exhausted pre training scaling and are looking to post training to expand capabilities, which is really just optimization via reinforcement learning against specific tasks aka bench maxing.
- ctolsen 16d agoTheir ambition isn't your work being amplified by their model, they want you running fifty autonomous long-running agents.