3 ms·
Reading Between the Dots: Decoding Hidden Computation Across Filler Tokens
- user_7832 2mo ago(To quote one of the authors) If you ask a frontier LLM a multi-hop reasoning question, e.g., "Who won the Nobel Prize for Chemistry in (1900 + Mozart's age when he died)?", it usually can't answer correctly immediately (no thinking) BUT if you ask the same question & append 300 dots, suddenly it can answer? (Because the LLM starts thinking "during" the dots).
- user_7832 2mo agoTweet: https://x.com/kaleybrauer/status/2078185882926846044 https://x.com/kaleybrauer/status/2078185882926846044 https://xcancel.com/kaleybrauer/status/2078185882926846044 https://xcancel.com/kaleybrauer/status/2078185882926846044 Explanatory image: https://x.com/kaleybrauer/status/2078185882926846044/photo/1 https://x.com/kaleybrauer/status/2078185882926846044/photo/1