Comment by xlbuttplug2
1 day ago
I presume post training is significantly easier than the distillation/training the top Chinese labs are doing.
I wonder if, similar to the American labs, they'll become stingy with their weights once they start getting immediately undercut by a wave of slightly better derived models.
No comments yet
Contribute on Hacker News ↗