Comment by snehesht
10 hours ago
Yeah I agree, I'm running it with Pi didn't notice much difference compared to lower tier models and the speed, of course.
10 hours ago
Yeah I agree, I'm running it with Pi didn't notice much difference compared to lower tier models and the speed, of course.
I am running 27B with Deepseek Harness these days and somehow just by using it, without any parameter changes, the model feels even more intelligent.
do LLMs tend to be homesick when not used in the same harness they sat in during some training phase?
iirc there was a sectionin Qwen’s paper where they talked anout how they post-trained flash or 3.8 to work just as well regardless of the harness or eval used. I think that used to be true but not sure if it is any longer