2 ms·
I know this is a popular position and it makes sense at face value when you think of LLMs as autocomplete systems. But it’s genuinely wrong. Relevant reading i
by scrollaway 17d ago
I know this is a popular position and it makes sense at face value when you think of LLMs as autocomplete systems. But it’s genuinely wrong.
Relevant reading is most notably anthropic’s research on the J-space. LLMs will plan ahead of time helped with CoT, get to a plan and “store” it in j-space, and execute on that plan which means they can in fact “backtrack” and give you reasoning on why they did something, because it IS part of their state.
- catlifeonmars 17d agoIs there a way to dump the state directly?
- scrollaway 16d agoNot unless you are anthropic/openai, or run your own models.
- orbital-decay 17d agoJ-space is just one convenient projection (out of many) to look at these well known phenomena