Comment by janalsncm
7 hours ago
Right, if a model says it is Qwen there is no way to distinguish a ModernBert fine tuned with Qwen completion data from a Qwen model fine tuned with completion data.
It’s also entirely possible that they used completions from a pool of open weight models.
No comments yet
Contribute on Hacker News ↗