3 ms·
When ChatGPT sees '473936482738338373926', it tokenizes it into 7 tokens, not 20. You can play with it here: https://platform.openai.com/tokenizer https://platf
by mquander 3y ago
When ChatGPT sees '473936482738338373926', it tokenizes it into 7 tokens, not 20. You can play with it here: https://platform.openai.com/tokenizer https://platform.openai.com/tokenizer
So to reverse the string, it can't just learn a kind of generic string reversal algorithm that goes character-by-character and outputs a result character-by-character. If it learned to output the tokens in reverse order, it would give a wrong answer (something like '926373...936473'.) Instead it has to learn something more complicated, like how each token is related to some other token that is its reverse, or how to identify it as a computational task and run some Python.
- oivey 3y agoRight, so it isn’t intuitively a hard task. It’s hard in the framework of ChatGPT’s design. That design is majorly powerful, but isn’t a great one for this basic task.
- arthurcolle 3y agowith a tool, an LLM agent can easily perform this task. If it evaluates its performance against a bunch of different tools and then learns, this is learnable. just not with pure LLM approach