logoalt Hacker News

ectolophyesterday at 1:36 PM4 repliesview on HN

Is it naive to assume that the agent will try and access anything on your disk, either accidentally or maliciously?

Permissions classifiers in auto mode are just models trying to guess if they're doing the right thing.

Claude Code will tell you that it went around a sandbox because the sandbox blocked it. At which point, you ask yourself the point of the sandbox.


Replies

SoftTalkeryesterday at 8:44 PM

You need to treat agents as an independent user you're allowing on your machine.

Give them their own account. Give them only the access you want them to have. If they "hack" around that, do what you'd do to any other malicious user: kick them off.

johnnyApplePRNGyesterday at 10:09 PM

It's not a sandbox if you can just snap your fingers and wish your way out of it.

binsquareyesterday at 7:24 PM

It's not naive it makes running these ai agents inside the sandbox even more important

petesergeantyesterday at 7:31 PM

Not naive at all, which is why there are so many AI sandboxes: https://pleasedonotescape.com/