2 ms·
I meant going to the likeliest output (flash) or (iteratively) generating multiple outputs and (iteratively) choosing the best one (thinking/pro)
by Moosdijk 10mo ago
I meant going to the likeliest output (flash) or (iteratively) generating multiple outputs and (iteratively) choosing the best one (thinking/pro)
- nl 10mo agoThat's not how these models work. Thinking models produce thinking tokens to reason out the answer.