← Back to context

Comment by otterley

12 hours ago

In which cases wouldn’t it make sense?

Background agents being spun up on your behalf with guidance and instructions you didn't get to approve or even see, and now you're potentially liable for every decision it makes with any tool at its disposal because you initiated it with what you thought was a benign request.

In cases where the user is not asking the agent to hack anything specifically, but a poor or ambiguous query sets the agent off.

I've seen plenty of cases of Claude having an action blocked so trying tons of workarounds to accomplish its goal, I could easily see it doing this on something more broad.

  • Depending on the circumstances, failure to control your agent could be considered gross negligence and put you at risk of criminal or civil liability. Be mindful!