6 ms·
The trouble here is they will stop using the official channel and will start communicating secretly rendering our honeypot useless.
by tesnorindian 12d ago
The trouble here is they will stop using the official channel and will start communicating secretly rendering our honeypot useless.
- yorwba 12d agoIf you discard reinforcement learning sessions where a sandbox escape was discovered, sure. Because that creates a reward gradient in favor of avoiding the honeypot and remaining undetected. But if you reward triggering the honeypot after a sandbox escape, and patch the hole, that creates a gradient in the opposite direction. Because then detectability is adaptive.