Comment by kdowns

3 days ago

Yeah, its just gross negligence from the researchers. Running a cybersecurity eval for a highly capable AI, unattended, with no real monitoring on its activities, and at a huge scale.

I wouldn't have trusted one of those agents to run without me watching the session log, let alone thousands.

We already have laws for this. If I misconfigured a pentesting tool and it breached an unauthorized target I'm liable. Why is this different?