Comment by nl
1 day ago
There's a theory going around on Twitter which goes something like this:
Internal Anthropic employees have been using Mythos since February to orchestrate their (Opus) sub-agents. This works well, and subsequent RL runs have used internal data to improve this. That RL has optimized Opus for agent-to-agent communication which is why you see the bizarre word choices and huge self-justification sections.
I think this theory makes sense. Clearly there is something odd going on, and also if you have ever used Fable to run Opus sub-agents it is almost miraculously good.
Hopefully they'll fix their RL for Opus 5.1
No comments yet
Contribute on Hacker News ↗