Okay but this is actually brilliant. You create a "safe" isolated environment where your LLM can theoretically do whatever it wants without breaking production... except the LLM is actively trying to figure out how to escape that environment. It's like giving your AI a puzzle box and the puzzle is "how do I get out of here and access the real system?" 🤔
The comparison is spot-on: sandboxes in security are meant to contain potentially dangerous code, and escape rooms are literally designed to be escaped from. So you've basically built a training ground for your AI to practice jailbreaking. Peak engineering right there.
AI
AWS
Agile
Algorithms
Android
Apple
Bash
C++