4 ms·
I don’t think text is sufficiently unconstrained. It is trivial to get GPT to output “Hi, how are you?” but of course it’s impossible to determine if that commo
by stu2b50 4y ago
I don’t think text is sufficiently unconstrained. It is trivial to get GPT to output “Hi, how are you?” but of course it’s impossible to determine if that common turn of phrase is human or machine generated.
There is too much a correlation between the position of tokens and correct meaning. In contrast, in an image, it’s very possible to have pixels shuffle around and not change the look. There, it is more likely that we can determine that the output is statistically more likely to be from a particular model.
- AlotOfReading 4y agoYeah, trivial text samples are obviously insufficient. At the other extreme (say millions of tokens of output), there's probably a reasonable chance that you could reliably distinguish human and LLM output. I suspect there might be some interesting questions down the rabbit hole in the middle though.