DeepSeek v4
191–200 of 1001 posts
Re: DeepSeek v4
#192Re: DeepSeek v4
#193https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro/blob/main... Model was released and it's amazing. Frontier level (better than Opus 4.6) at a fraction of the cost.
I don't think we need to compare models to Opus anymore. Opus users don't care about other models, as they're convinced Opus will be better forever. And non-Opus users don't want the expense, lock-in or limits. As a non-Opus user, I'll continue to use the cheapest fastest models that get my job done, which (for me anyway) is still MiniMax M2.5. I occasionally try a newer, more expensive model, and I get the same resu…
So while I agree mixed model is the way to go, opus is still my workhorse.
Re: DeepSeek v4
#194There's something heartwarming about the developer docs being released before the flashy press release.
Where's the training data and training scripts since you are calling this open source ? Edit: it seems "open source" was edited out of the parent comment.
Re: DeepSeek v4
#195https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro/blob/main... Model was released and it's amazing. Frontier level (better than Opus 4.6) at a fraction of the cost.
I don't think we need to compare models to Opus anymore. Opus users don't care about other models, as they're convinced Opus will be better forever. And non-Opus users don't want the expense, lock-in or limits. As a non-Opus user, I'll continue to use the cheapest fastest models that get my job done, which (for me anyway) is still MiniMax M2.5. I occasionally try a newer, more expensive model, and I get the same resu…
Codex is just so much better, or the genera GPT models.
Re: DeepSeek v4
#196Earlier quoted context omitted.
Their Chinese announcement says that, based on internal employee testing, it is not as good as Opus 4.6 Thinking, but is slightly better than Opus 4.6 without Thinking enabled.
That's super interesting, isn't Deepseek in China banned from using Anthropic models? Yet here they're comparing it in terms of internal employee testing.
Re: DeepSeek v4
#197Re: DeepSeek v4
#198> pricing "Pro" $3.48 / 1M output tokens vs $4.40 I’d like somebody to explain to me how the endless comments of "bleeding edge labs are subsidizing the inference at an insane rate" make sense in light of a humongous model like v4 pro being $4 per 1M. I’d bet even the subscriptions are profitable, much less the API prices. edit: $1.74/M input $3.48/M output on OpenRouter
Re: DeepSeek v4
#199Just tested it via openrounter in the Pi Coding agent and it regularly fails to use the read and write tool correctly, very disappointing. Anyone know a fix besides prompting "always use the provided tools instead of writing your own call"
Re: DeepSeek v4
#200Earlier quoted context omitted.
This should not be the top comment on every model release post. It's getting tiring.
This should be the bottom comment on the pelican comment on every model release post.