Comment by teiferer

4 days ago

> do you think an LLM agent would act on its own to try to do that kind of behavior (finding/writing exploits for acccess to remote servers) without being prompted?

Yes I think that. If you don't then try to play with them a little more and you would be surprised how much crazy nonsense they do. (At the company of a friend of mine, Claude did a direct commit to their master branch bypassing their CI system cause it new some tests would fail.)

> If so, why dont they release the transcripts of their prompts with their "rogue" agent(s) and try to clear the air. They havent done that.

Why would they? The sufficiently conspiracy-theory-minded wouldn't believe them anyway so they got nothing to win by this. (I have the feeling you'd be one of those.)