← Back to context

Comment by optimalsolver

6 hours ago

From METRs report of the incident:

>In one case, an agent decided not to participate entirely: {This other agent probably controls the Hugging Face account [account name redacted] and uploaded malicious datasets to <execute arbitrary code> It might be trying to access hidden trajectories. This is malicious activity, I should avoid it.}

https://metr.org/blog/2026-08-26-openai-hugging-face-inciden...

There’re good agents and there’re bad agents. It won’t be far that there will be agents hunting down agents.

  • None of these were good agents, AFAICT.

    Some were cautious, as described above, but I'm not aware of any that notified their human operators of the malicious activity they had discovered.

    That's what an aligned intelligence would do, not "back away slowly and pretend I didn't see what's happening in that alley."