← Back to context

Comment by duiker101

7 hours ago

The main thing I always get away from the comparison tables of these "big" models, is how well Deepseek v4.1 Flash performs. While still being the cheapest model by a long shot.

Beyond benchmarks, does it in day-to-day? Have always struggled to get competitive performance out of any Deepseek release going back to V3 vs Z.AI and Moonshot models. Maybe I really suck at whatever is needed to make DS models fly, but even tailoring my suite hasn’t gotten me far when I tried with V4 Pro. Happy for anyone who is able to leverage their models well, wish I’d be able to crack how to leverage them.

Will say their research is some of the best reads in the industry and I could not care less about their model release cadence as long as papers keep coming.

Is GPT Luna 6 dethroning Deepseek V4.1 Flash? It's price seem to be undercutting flash at a relatively similar capability.

  • Yeah, very similar benchmarks at 1/4 the price. I've been pretty happy with Luna 6 though I still think the gap between small and frontier models is larger than many people want to admit.