3 ms·
Is this motivated by the value of the thinking traces gleaned from the traffic?
by josh-wrale 2mo ago
Is this motivated by the value of the thinking traces gleaned from the traffic?
- ec109685 2mo agoThey can’t decrypt the thinking traces.
- dannyw 2mo agoYou can train a LLM to inverse summarised thinking into thinking text. It’s not perfect, but it gets you maybe 80% of the quality with proper techniques. Paper: https://arxiv.org/abs/2603.07267 https://arxiv.org/abs/2603.07267 FWIW, there’s not that much value protected here anyway IMHO, and even raw thinking text can lie (as shown by Anthropic’s amazing research), so for legitimate interpretability research it’s limited. Scaling frontier performance hasn’t been SFT-bounded for a while now; it’s now basically how much you can scale RL rollouts.
- killingtime74 2mo agoThe thinking traces are server-side, not exposed