It’s four poorly constructed arbitrary experiments which say very little about the competency of either model. The article reads like thin, auto-generated ai clickbait for nerd sniping or shilling a model. Consider the lead: > DeepSeek V4 Pro wins this head-to-head by being more exact where it matters: following instructions, matching schemas, and solving edge cases cleanly. GPT-5.5 Pro is still strong, but it gave a…
DeepSeek V4 Pro beats GPT-5.5 Pro on precision
211–220 of 249 posts
Re: DeepSeek V4 Pro beats GPT-5.5 Pro on precision
#212I've been using deepseek v4 for cost/performance reasons. I feel it is generally not as good as some others, but in the end, you can make any model work by giving it the right acceptance criteria. Use detailed specs, use tests, and give it the power to iterate until it works. One-shot is a poor metric for performance.
It helps but you often have to step in the failure cases and guide them or forcibly fix certain paths to get a solution.
Re: DeepSeek V4 Pro beats GPT-5.5 Pro on precision
#213Curious for folks who have made the switch I’m considering: if I swapped Claude Code to DeepSeek API pricing, would I get more bang for my buck compared to the $100 Max plan I’m using now? I only hit the 5 hour limit every few days and the weekly limit a day or two before it resets at the most aggressive. I wouldn’t expect my usage to increase dramatically, other than not being stopped by limits. I’m still apprehensi…
DeepSeek and Xiaomi's deals on cache reads go with their models' latest gens making caching cheaper (using less space for KVs). No open-model inference provider has decided to match the pricing. I'm sure that says something about how inference pricing works, but not completely sure what.
Agree with others that top open models aren't on the frontier, and I would expect differences doing big-picture planning or anywhere you're only giving broad brushstrokes and looking for a lot to be guessed. But they do seem fine at coding from a a concrete plan! No experience in huge codebases because I only use them outside work, but they seem good enough about gathering info before they dive in that I'd expect them to grep around as they need.
An annoying caveat: individual subscription plans, used heavily, are much cheaper than the API -- see https://she-llac.com/claude-limits -- which complicates any argument about cost. I still think open models are worth playing with. They're one of the things that let us treat this as a technology rather than just as the product offerings of one of a few companies.
Re: DeepSeek V4 Pro beats GPT-5.5 Pro on precision
#214It’s four poorly constructed arbitrary experiments which say very little about the competency of either model. The article reads like thin, auto-generated ai clickbait for nerd sniping or shilling a model. Consider the lead: > DeepSeek V4 Pro wins this head-to-head by being more exact where it matters: following instructions, matching schemas, and solving edge cases cleanly. GPT-5.5 Pro is still strong, but it gave a…
> poorly constructed arbitrary experiments which say very little about the competency of either model. No one ever says this about the “pelican on a bicycle” metric
Re: DeepSeek V4 Pro beats GPT-5.5 Pro on precision
#215Earlier quoted context omitted.
Yeah, the discounted deepseek inference is subsidized by the CCP for a reason, and it's one that might well come back to bite.
> deepseek inference is subsidized by the CCP What is that claim based on?
[1] https://chinaselectcommittee.house.gov/sites/evo-subsites/se... [2] https://ai.americansecurityproject.org/news/ai-imperative-20...
and more.
Of course, you can choose to ignore America-biased sources, but since it aligns with the obvious.
Re: DeepSeek V4 Pro beats GPT-5.5 Pro on precision
#216Earlier quoted context omitted.
Yeah, the discounted deepseek inference is subsidized by the CCP for a reason, and it's one that might well come back to bite.
There is no evidence it is subsidized. Actually, there is evidence that (1) electricity is cheap in China & (2) deepseek is a very efficient model.
Re: DeepSeek V4 Pro beats GPT-5.5 Pro on precision
#217Earlier quoted context omitted.
There are monied interests that do not want inexpensive Chinese successors to Scam Altman's creation.
They're inexpensive because they're derived from his creation.
Re: DeepSeek V4 Pro beats GPT-5.5 Pro on precision
#218Earlier quoted context omitted.
> deepseek inference is subsidized by the CCP What is that claim based on?
Besides common sense given the clear geopolitical context, sources like: [1] https://chinaselectcommittee.house.gov/sites/evo-subsites/se... [2] https://ai.americansecurityproject.org/news/ai-imperative-20... and more. Of course, you can choose to ignore America-biased sources, but since it aligns with the obvious.
*This does not invalidate other concerns (censorship, privacy) but the way people phrase it makes it look like DeepSeek and co. are 'cheating' somehow with their business model by 'distorting' inference cost to make it way artificially lower than its 'natural price' (either notion being hopelessly naive)
Re: DeepSeek V4 Pro beats GPT-5.5 Pro on precision
#219Earlier quoted context omitted.
I think you've misunderstood the purpose of a lead (sic). Per Merriam-Webster [^1], a lede is: > the introductory section of a news story that is intended to entice the reader to read the full story (Emphasis mine) You may prefer more matter-of-fact phrasing, of course, but criticising a lede for attempting to achieve its goal is unjustified. [^1]: https://www.merriam-webster.com/dictionary/lede
A 'lede' is just an intentionally differentiated spelling of 'lead'; the origin of the word is just lead . Collins dictionary defines lede: a variant spelling of lead