Comment by postalcoder
1 day ago
I hope DeepSeek takes some time to improve their tuning for reasoning effort. Right now, there are only three reasoning efforts: low, high, and max.
For all intents and purposes, "low" is pretty much the same as turning reasoning off, and "high" is similar to "max". "High/max" performs way too much reasoning, takes forever, and causes costs to balloon. They need a proper "medium" setting.
I get it that they're probably focused on pushing performance right now, but the ergonomics of the model aren't great.
I switched to GLM-5.3 flash on high for this reason. Too many "but wait" in the Deepseek-v4 reasoning.
GLM 5.3 Flash was the killer for me. It really feels like we have Claude-approaching models at home.
I just wish they kept parameter count down in order to fit entirely within commonly used RAM sizes
I would rather have just 3 levels: low, medium and high.