Comment by daft_pink 3 months ago Is this true with Groq too? 3 comments daft_pink Reply siwakotisaurav 3 months ago Groq doesn’t have r1, only a llama 70b distilled with r1 outputs. Kinda crazy how they just advertise it as actual r1 daft_pink 3 months ago I don’t quite understand what the difference between the Groq version and the actual r1 version are. Do you have a link or source that explains this? Alex-Programs 3 months ago Actual r1 is a base model by Deepseek with CoT added by them via RL.The distillations are other people's base models with RL used to add CoT to them. The 70b one is Llama that Deepseek has modified.As I understand it.
siwakotisaurav 3 months ago Groq doesn’t have r1, only a llama 70b distilled with r1 outputs. Kinda crazy how they just advertise it as actual r1 daft_pink 3 months ago I don’t quite understand what the difference between the Groq version and the actual r1 version are. Do you have a link or source that explains this? Alex-Programs 3 months ago Actual r1 is a base model by Deepseek with CoT added by them via RL.The distillations are other people's base models with RL used to add CoT to them. The 70b one is Llama that Deepseek has modified.As I understand it.
daft_pink 3 months ago I don’t quite understand what the difference between the Groq version and the actual r1 version are. Do you have a link or source that explains this? Alex-Programs 3 months ago Actual r1 is a base model by Deepseek with CoT added by them via RL.The distillations are other people's base models with RL used to add CoT to them. The 70b one is Llama that Deepseek has modified.As I understand it.
Alex-Programs 3 months ago Actual r1 is a base model by Deepseek with CoT added by them via RL.The distillations are other people's base models with RL used to add CoT to them. The 70b one is Llama that Deepseek has modified.As I understand it.
Groq doesn’t have r1, only a llama 70b distilled with r1 outputs. Kinda crazy how they just advertise it as actual r1
I don’t quite understand what the difference between the Groq version and the actual r1 version are. Do you have a link or source that explains this?
Actual r1 is a base model by Deepseek with CoT added by them via RL.
The distillations are other people's base models with RL used to add CoT to them. The 70b one is Llama that Deepseek has modified.
As I understand it.