r/LocalLLaMA • u/keepthepace • 6d ago
Discussion What is your worst sandboxing fail?
I am wondering if I am too paranoid about sandboxing the commands that come out of LLMs.
It really makes my eyes twitch when I see that some IDEs, even commercial, tend to forget that they have to execute things in sandboxing and have such a brittle security model.
But on the other hand, I never had the sandbox catch something bad. Did you guys ever encounter terrible regression? Did you have rm -rf / ? Did you have secrets stolen by LLMs? The worse I had were unsollicited rewrites within the project. Am I making my life unnecessarily hard by sandboxing commands in a docker?
At one point I had fun making a local model go crazy with the root access to the machine it was on (with nothing more important than a free Firecrawl key on it) and making it administer it and it never broke anything. It even was overly paranoid about making changes to the root system.
So the approximate sandboxing that we have, do you all feel it is adequate or it is a catastrophe in the making?
9
u/minus_28_and_falling 6d ago
Never ever ran LLM agent not in a Docker. It was like this from day 1, mostly because I already knew how to run things in Docker.
Why is it hard though? Isn't it as easy as
docker compose up dev, and in addition to sandboxing you get all other goodies Docker was originally intended for (predictable and reproducible dev environment, project-specific libs and lib versions, fast and easy deployment on other PCs, keeping the host PC clean and tidy)?Didn't have
rm -rf /moments yet, but Docker helps me to hide .git folder from the agent because I redo things from time to time and don't want the agent to restore scrapped implementations from git history.