Comment by walrus01
6 hours ago
I would think it's entirely plausible that they have so many R&D agents/LLMs in active use at any one time that it's far beyond the capacity of any human to review the log files of their activity. Even just to go through the reasoning. It's hard enough for 1 person running opencode to keep up with the reasoning from 1 very verbose/long-thinking LLM with fast tok/s output for a small discrete single-purpose project.
Whatever OpenAI is doing, if it's being properly logged, it must be a firehose of logs.
>it's far beyond the capacity of any human to review the log files of their activity
Maybe they should contract with one of the other AI labs. I hear they have LLMs that are good at that kind of thing.