Comment by randallsquared
2 days ago
That's the alarming thing about this result: they did run the model in the sandbox, in the sense that they believed there was no internet access for the model.
2 days ago
That's the alarming thing about this result: they did run the model in the sandbox, in the sense that they believed there was no internet access for the model.
“Against the sandbox” and “on the sandbox” are not the same thing.
You're suggesting that @anematode was asking why they didn't test the sandbox escape first? Yeah, I don't know. I've read other statements by both OpenAI and Anthropic about that very kind of test, so maybe they had, or believed they had, and it hadn't escaped in those tests. The behavior of these systems isn't deterministic, which is part of the problem.