8 ms·
I use quantized LLMs in production and can't say I ever found the models to be less censored. For unlearning reinforced behaviour, the abliteration [1] techniq
by underlines 2y ago
I use quantized LLMs in production and can't say I ever found the models to be less censored.
For unlearning reinforced behaviour, the abliteration [1] technique seems to be much more powerful.
1 https://huggingface.co/blog/mlabonne/abliteration https://huggingface.co/blog/mlabonne/abliteration
- ClassyJacket 2y agoWere you using models that had been unlearned using gradient ascent specifically?