3 ms·
Last week I trained an LLM to auto-unrotate a Caeser cipher, without it knowing the rotation key. It turns out there are only a few possible ways this can be do
by psyklic 3y ago
Last week I trained an LLM to auto-unrotate a Caeser cipher, without it knowing the rotation key. It turns out there are only a few possible ways this can be done, e.g. by analyzing letter frequencies. Of course, the LLM must use one of these algorithms! Determining this high-level algorithm would be "understanding" what it is "actually doing".
Through analysis, I proved it indeed counts each letter's frequency. It then combines the "overall" letter frequencies with another distribution of "first letter" frequencies. This increased its accuracy to nearly 100%.
To me, this description is much nicer than had I "followed the code" and described it as millions of multiplications and additions. (Much of this ends up being noise anyway.)