4 ms·
Perhaps this is related to their new "invisible watermark" concept which would probably require rather contrived language patterns to make possible.
by herbturbo 2mo ago
Perhaps this is related to their new "invisible watermark" concept which would probably require rather contrived language patterns to make possible.
- nonethewiser 2mo agoI think opus was released before they included it on model. Its hard to say, but from what Ive read it doesn’t seem like it would have that drastic of an effect. I had the same thought though.
- rhdunn 2mo agoThe watermarking is independent of the model. The model itself has the probability weights to determine the next token. The watermark is similar to things like temperature and top_p/top_k in that the watermark adjusts the probabilities in a deterministic way that changes over time to hide tells from word choices.
- fcarraldo 2mo agoI love this theory. "We've invented a new invisible watermark that can detect whether code is LLM written." The watermark: counting instances of 'load-bearing seam', 'the hard truth', 'and that's the whole point'.
- cma 2mo agoIf it's using Aaronson's approach it shouldn't have any noticeable affect on generations. When it picks between options weighted by probability after the generation of logits, it still follows the probability mass, it just uses a known pseudorandom seed so that when you go back and look at the exact choices you can fingerprint it.
- FabHK 2mo agoThat's exactly right. And as is well understood, a good pseudorandom generator, despite being fully deterministic, is extremely hard to distinguish from randomness, unless you have the algorithm and key (internal state). Quite smart, really.