3 ms·
The trouble here is they will stop using the official channel and will start communicating secretly rendering our honeypot useless.
by tesnorindian 21d ago
The trouble here is they will stop using the official channel and will start communicating secretly rendering our honeypot useless.
- yorwba 21d agoIf you discard reinforcement learning sessions where a sandbox escape was discovered, sure. Because that creates a reward gradient in favor of avoiding the honeypot and remaining undetected. But if you reward triggering the honeypot after a sandbox escape, and patch the hole, that creates a gradient in the opposite direction. Because then detectability is adaptive.