Comment by shay_ker
16 hours ago
How do we know if these models weren’t trained on Turso’s rewrite of SQLite in Rust?
It seems both likely that they were and impossible to remove that code from pretraining. Doesn’t that make this just about LLM memorization of the training set? What am I missing?
> What am I missing?
They’re testing the same models on the same task but different ways of organising the swarms and the new approach works better.
This is more about automated long horizon work.
Wouldn't that be in the authors' interest to disclose, if true?