3 ms·
I'd be extremely surprised if AI labs are not doing or planning on doing this already. The same way that reasoning models are trained on chain of thoughts, why
by silveraxe93 2y ago
I'd be extremely surprised if AI labs are not doing or planning on doing this already.
The same way that reasoning models are trained on chain of thoughts, why not do it with program state?
Just have a "separate" scratchpad where the AI keeps the expected state of the program. You can verify if that is correct or not. Just use RL to train the AI to always have that correct.
- deleted 2y ago[deleted]