Comment by daft_pink
3 months ago
I don’t quite understand what the difference between the Groq version and the actual r1 version are. Do you have a link or source that explains this?
3 months ago
I don’t quite understand what the difference between the Groq version and the actual r1 version are. Do you have a link or source that explains this?
Actual r1 is a base model by Deepseek with CoT added by them via RL.
The distillations are other people's base models with RL used to add CoT to them. The 70b one is Llama that Deepseek has modified.
As I understand it.