3 ms·
They train the model, then as soon as they get their numbers, they let the safety people RLHF it to death.
by aedon 3y ago
They train the model, then as soon as they get their numbers, they let the safety people RLHF it to death.
- sebzim4500 3y agoI think it's just really hard to assess the performance of LLMs. Also AI safety is the stated reason for Anthropic's existence, we can't be angry at them for making it a priority.