3 ms·
Very cool! One question, is this model gimped with safety "features"?
by asdasdddddasd 3y ago
Very cool! One question, is this model gimped with safety "features"?
- flangola7 3y agoI don't know what you mean by "gimped", but they do advertise that it has safety and capability features comparable to OpenAI models, as rated by human testers.
- seydor 3y agoapart from the non-chat model, there are 2 chat models: > Others have found that helpfulness and safety sometimes trade off (Bai et al., 2022a), which can make it challenging for a single reward model to perform well on both. To address this, we train two separate reward models, one optimized for helpfulness (referred to as Helpfulness RM) and another for safety (Safety RM)
- logicchains 3y agoThe LLaMA chat model is, the base model is not.