logoalt Hacker News

user43928today at 9:26 AM1 replyview on HN

I doubt that it is more efficient for someone to routinely watch every line of output or stop to review terminal commands, rather than waiting for the turn to complete.

It is a broader debate about agentic AI, and whether one should relinquish control to the tool rather than aim for full understanding of every action taken.

The people arguing for a hands-on, fully in control approach are losing ground by the week, in my opinion.


Replies

mcmcmctoday at 4:20 PM

It’s definitely on the high friction side of the usability/security tradeoff, I just think auto modes like this are as liability-inducing as handing a script kiddie intern full admin on your production environment

The real answer is somewhere in the middle and is probably a mix of traditional AV/EDR and AI QA judges that mitigate risk of running more or less random arbitrary code and auto approve based on configured detection rules and your personal risk tolerance. Would it suck to stick an EDR sensor in every code execution environment spun up for an agent to run a python script… yes. It would also suck if you were responsible for hacking a company without knowing about it because you didn’t watch what your AI was doing