[flagged]
DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
181–190 of 342 posts
Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#182> For the Code Agent tasks among the public benchmarks above, DeepSeek-V4-Flash-0731 is evaluated with the minimal mode of DeepSeek Harness (to be released) as the agent framework So, are they planning to announce an optimized coding agent harness as well ? DSv4 flash is a fantastic model, and my daily driver. With reasonix or pi, I can code all day long and pay a few pennies for it. No token anxiety. Whereas the sam…
Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#183So GLM 5.2/Gemini 3.6 level intelligence for $0.28/m output. And their updated Pro model coming soon.... Plus a size you can genuinely run at home: Unsloth lossless Q8 at 162GB.
I would like to see your "home"
But it makes little to no sense as long as API prices are what they are. Except for maybe privacy reasons.
Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#184Already beat Luna on price/task, by about 2x: https://artificialanalysis.ai/models/deepseek-v4-flash?intel...
Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#185Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#186Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#187Earlier quoted context omitted.
Because reddit unironically has better decorum around usage of their upvote/downvote system than HN does. People on HN downvote objectively correct information because they don't like it 24/7. There's a reason the creator of Zig left and gave the computer version of a middle finger on the way out to HN!
fwiw pg said early on that downvoting for disagreement is perfectly fine: https://news.ycombinator.com/item?id=117171 commenting about voting is also something the HN guidelines warns against: > Please don't comment about the voting on comments. It never does any good, and it makes boring reading. https://news.ycombinator.com/newsguidelines.html
Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#188Earlier quoted context omitted.
Philosophically, I don’t believe we should outsource the accuracy of scripture to any single entity (let alone a for profit secular one). So it’s less about model choice and more about governance of scripture. I will check out the link you sent for sure!
I was under the impression you were getting BOOK.CHAPTER.VERSE references from the model and then sourcing them from a ground truth db. Anyways, impressive app! We haven't tackled such an ambitious project just for it being daunting.
Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#189> For the Code Agent tasks among the public benchmarks above, DeepSeek-V4-Flash-0731 is evaluated with the minimal mode of DeepSeek Harness (to be released) as the agent framework So, are they planning to announce an optimized coding agent harness as well ? DSv4 flash is a fantastic model, and my daily driver. With reasonix or pi, I can code all day long and pay a few pennies for it. No token anxiety. Whereas the sam…
They announced it already, read the tech report. "For the Code Agent tasks among the public benchmarks above, DeepSeek-V4-Flash-0731 is evaluated with the minimal mode of DeepSeek Harness (to be released) as the agent framework , using the max reasoning effort level with temperature = 1.0, top_p = 0.95."
I meant something I could download and run.
Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#190Earlier quoted context omitted.
Good luck when the last non-Chinese frontier labs will have closed and the CCP will ask to stop sharing models open source.
My conspiracy theory is that this is the new space race, and the CCP encourages this to show the world what Chinese engineers are capable of, and tank the Anthropic/OpenAI valuation bubble as a desirable side effect.