Live data from Hacker News

DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis

artificialanalysis.ai

11–20 of 342 posts

Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis

#14
post #12

Does it already know the answer to what happen at Tiananmen Square? Or still avoiding it?

Who cares if it is programming correctly I would be more worried about it not doing things like find security bugs because US or Chinese government does not want to. Which LLM is more likely to do that?

Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis

#15
New Deepseek models are like Christmas for me. Really big fan of low cost API models, noone does it better than DS. Until VRAM price is low enough to run models locally, this is the way to go.

The subsidized subscription model won't last, API pricing "feels" closer to a true sustainable business model.

Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis

#17
post #8

Already beat Luna on price/task, by about 2x: https://artificialanalysis.ai/models/deepseek-v4-flash?intel...

Maybe I'm reading that incorrectly, but it seems to me the cost is on the X-axis. First, your direct comparison, Deepseek V4 Flash 0731 (max effort) $0.03 (rounded up) per task @ index 50. OpenAI Luna: * high effort $0.03 (rounded down) @ index 46 * xhigh effort $0.04 @ index 49 * max effort $0.07 @ index 51 So I would say a fair statement would be "OpenAI Luna between 2x and 3x the price of Deepseek Flash, what you…

For anything substantial, you'd want a bigger model anyway.

For simple tasks, they're already saturated, and you'd prefer the faster model, so that you can have a realtime/interactive-ish experience.

Or to put it bluntly, it's cheaper if you don't value your time. That goes for smaller models in general -- need more handholding, more correcting -- but the Chinese ones are slower on top of that.

As for speed, Sol on Low is faster than Luna on most settings.

Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis

#19
post #12

Does it already know the answer to what happen at Tiananmen Square? Or still avoiding it?

It’s open weight, you can (or you can wait for someone else to) uncensor it. We shouldn’t be upset at the researchers making this for the mandates their government puts on them.

Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis

#20

If deepseek v4 flash is beating DeepSeek V4 Pro, can we expect new V4 Pro which is on par with Opus 5 in couple weeks (even better if it beats Opus)?

I’ve been using v4 flash for an app I’m building [1] and it’s amazing how cost effective and good it is coming from having always used gpt, opus and sonnet models.

It’s so cost effective I can offer a generous free tier since my goal isn’t to make money with it.

[1] https://trysojourn.app

Post reply on HN