Comment by ssalbiz
1 month ago
My understanding is that RLVR, synthetic data generation and a slew of other post-training techniques are what have driven many recent advances in models more so than manual data providers. The economics of that are for sure worse than just scaling pre-training but it is incorrect to think that test time inference scaling and manual data entry are the only ways in which models are advancing.
No comments yet
Contribute on Hacker News ↗