← Back to context

Comment by Conol_ai

11 hours ago

It's also worth reading this next to the Hugging Face incident from July and the METR report on it: HF was framed as a boundary failure during an eval. The question for every security team is now uncomfortable: does your access logging even distinguish agent behavior from ordinary automation, or would you find your own incident the way Australia apparently did — from the vendor, months late?