4 ms·
And that research is valuable, but it's definitely nowhere near the point where safety/alignment progress can't be made without a user base. A few recent favori
by jimrandomh 4y ago
And that research is valuable, but it's definitely nowhere near the point where safety/alignment progress can't be made without a user base. A few recent favorites of mine: https://arxiv.org/pdf/2212.03827.pdf https://arxiv.org/pdf/2212.03827.pdf https://www.alignmentforum.org/posts/aPeJE8bSo6rAFoLqg/solidgoldmagikarp-plus-prompt-generation https://www.alignmentforum.org/posts/aPeJE8bSo6rAFoLqg/solid...
- the8472 4y agoNote that that kind of research was only possible on the (smaller) open models. When the new models are proprietary ones then outsiders can't do that kind of analysis. If they had regenerated the tokenization dictionary to match the training data instead of reusing an older set then this wouldn't transfer from GPT2 to ChatGPT.