3 ms·
Thank you. Some variation of Circuit Tracing[1] is what I was thinking of. It would be useful to visualize the tokens effect on inference in real time, and twea
by squircle 2y ago
Thank you. Some variation of Circuit Tracing[1] is what I was thinking of. It would be useful to visualize the tokens effect on inference in real time, and tweak the input parameters, individual tensor weights, etc., and watch the inference collapse on a different result (does output of LLMs need to be next token based? Would larger, pre trained/synthesized tokens with known and expected output allow for faster inference, or is this simply tool calling?)
I understand LLMs and ML models are not linear functions but, they are functions (no?) and some complex higher level math can be grasped intuitively when visualized.
[1] https://news.ycombinator.com/item?id=43495585 https://news.ycombinator.com/item?id=43495585
Edit: thoughts.