3 ms·
I would guess they are trying to maximize training data
by cousinbryce 5mo ago
I would guess they are trying to maximize training data
- Zak 5mo agoIf I was being rewarded for using more tokens, I would feed LLM output back into the model. That's probably not very useful training data.
- piva00 5mo agoI personally know two people who are doing exactly that after a mandate rolled out at their work, the measurement is "tokens spent" and since they weren't finding many cases that required a lot of tokens they simply started to run agent loops feeding each other. Absurdly wasteful but Goodhart's Law almost never fails.