4 ms·
It's not the training dataset. All of these models, including the "open" ones, have been RLHF'ed by teams of politically-motivated people to be "safe" after in
by caeril 2y ago
It's not the training dataset.
All of these models, including the "open" ones, have been RLHF'ed by teams of politically-motivated people to be "safe" after initial foundation training.
- pjkundert 2y agoAnd I’m not even remotely interested in the “corrections” supplied by some group of right-thinking meddlers! This corruption must be disclosed as assiduously as the base dataset, if not more so.
- _yid9 2y agoOr, at least package them up as "personnas" and give them an appropriate name, eg. "Church Lady", "Jr. Marxist Barista", "Undergrad Philosophy Major", ... Actually, those seem like an apt composite description of the PoV of the typical mass-market AI... 8/
- Der_Einzige 2y agoNot mistrals. Mistral large is willing to tell me how to genocide minorities or NSFW without any kind of orthogonalization or fine tuning. Please actually try models instead of pontificating without evidence. Try it for yourself: https://huggingface.co/mistralai/Mistral-Large-Instruct-2407 https://huggingface.co/mistralai/Mistral-Large-Instruct-2407
- pjkundert 2y agoI wasn’t aware that there was any publicly accessible interface to the Mistrals (or any other) models without training-wheels!