Comment by bmorg
3 hours ago
In-prompt "security" is not reliable. You can not tell if the LLM/agent actually followed your instructions or whether it fell for a prompt injection.
3 hours ago
In-prompt "security" is not reliable. You can not tell if the LLM/agent actually followed your instructions or whether it fell for a prompt injection.
No comments yet
Contribute on Hacker News ↗