Comment by nickandbro

16 hours ago

I wouldn't doubt GPT 4.6 Luna being in the top left quadrant's center on the Cost per Intelligence Index is not concerning for Liang Wenfeng. You have to remember DeepSeek v4 flash even though a bit cheaper, does not have vision abilities, which is a big draw for agentic tasks.

I admire DeepSeek's openness, but even they have been raising prices after their discounts.

They haven't raised prices though the plan was to increase it with peak hour usage for V4 Pro GA release, they didn't do that, so no price increases there I believe.

As for vision yeah it sucks but Luna is also 2x input and 1.5x output for 1M context...

That's around 0.4 in/1.8 out

DSv4 is wayyy cheaper.

And it's open now you have Luna at home if you have a decent set of GPUs you can run this on 2Sparks or one very expensive Mac or just like 6-8 5090s..

  • I know most folks can't afford it I am working on making it viable to rent shared hosting the biggest issue is data leak and prompt injection attacks with shared hosting. (Since the server owner connects to your main system via the coding agent)

    I guess using a ZDR provider is good enough for now.

According to the leaked call transcript, DeepSeek is working on vision for V4. Not sure when it will land though.

Gpt 5.6 Luna cache read is $0.02 per mtok

V4 flash cache read is $0.0028 per mtok

That's not "a bit cheaper", just saying

> I wouldn't doubt GPT 4.6 Luna being in the top left quadrant's center on the Cost per Intelligence Index is not concerning for Liang Wenfeng.

The leaked interview has him saying it doesn't matter... as much as open source doesn't matter. There's enough in it for everyone right now and they aren't after everything.

Perspective: DeepSeek doesn't have enough infrastructure to serve their target customers already.