Oh well, I should have bought 2x 512GB RAM MacStudios, not just one :(
DeepSeek v4
591–600 of 1001 posts
Re: DeepSeek v4
#592Just tested it via openrounter in the Pi Coding agent and it regularly fails to use the read and write tool correctly, very disappointing. Anyone know a fix besides prompting "always use the provided tools instead of writing your own call"
Re: DeepSeek v4
#593Earlier quoted context omitted.
Still not sure how I feel about China of all places to control the only alternative AI stack, but I guess it's better than leaving everything to the US alone. If China ever feels emboldened enough to go for Taiwan and the US descends into complete chaos, the rest of the world running on AI will be at the mercy of authoritarian regimes. At the very least you can be sure noone is in this for the good of the people anym…
Isn’t Mistral close in the ballpark?
I think they are leaders in the democratization of LLMs. Almost everyone has a computer right now that can run a useful variant of a Mistral model. I hope they keep their focus because what they are aiming for likely has the biggest impact on the average person and would be the best case scenario for the technology in general.
Re: DeepSeek v4
#594The incredible arrogance and hybris of the American initiated tech war - it is just a beautiful thing to see it slowly fall apart. The US-China contest aside - it is in the application layer llms will show their value. There the field, with llm commoditization and no clear monopolies, is wide open. There was a point in time where it looked like llms would the domain of a single well guarded monopoly - that would have…
I've been baffled watching America double down on the same strategy even when it failed to produce results They sanctioned the hell out of Huawei and now Huawei is bigger than ever America is just not able to digest the idea that another country can be as good, if not better, at innovation
Re: DeepSeek v4
#595Open Source as it gets in this space, top notch developer documentation, and prices insanely low, while delivering frontier model capabilities. So basically, this is from hackers to hackers. Loving it! Also, note that there's zero CUDA dependency. It runs entirely on Huawei chips. In other words, Chinese ecosystem has delivered a complete AI stack. Like it or not, that's a big news. But what's there not to like when…
As a Brit I'm here for it to be honest, I'm tired of America with everything that's going on. China is not perfect but a bit of competition is healthy and needed
Re: DeepSeek v4
#596The incredible arrogance and hybris of the American initiated tech war - it is just a beautiful thing to see it slowly fall apart. The US-China contest aside - it is in the application layer llms will show their value. There the field, with llm commoditization and no clear monopolies, is wide open. There was a point in time where it looked like llms would the domain of a single well guarded monopoly - that would have…
I've been baffled watching America double down on the same strategy even when it failed to produce results They sanctioned the hell out of Huawei and now Huawei is bigger than ever America is just not able to digest the idea that another country can be as good, if not better, at innovation
Re: DeepSeek v4
#597In my tests too[0], it doesn't reach top 10. One issue, which they also mentioned in their post, is that they can't really serve well the model at the moment, so V4-Pro is heavily rate-limited and gives a lot of timeout errors when I try to test it. This shouldn't be an issue though, considering the model is open-source, but it makes it hard to accurately test at the moment.
[0]: https://aibenchy.com/compare/deepseek-deepseek-v4-flash-high...
Re: DeepSeek v4
#598> pricing "Pro" $3.48 / 1M output tokens vs $4.40 I’d like somebody to explain to me how the endless comments of "bleeding edge labs are subsidizing the inference at an insane rate" make sense in light of a humongous model like v4 pro being $4 per 1M. I’d bet even the subscriptions are profitable, much less the API prices. edit: $1.74/M input $3.48/M output on OpenRouter
Re: DeepSeek v4
#599Open Source as it gets in this space, top notch developer documentation, and prices insanely low, while delivering frontier model capabilities. So basically, this is from hackers to hackers. Loving it! Also, note that there's zero CUDA dependency. It runs entirely on Huawei chips. In other words, Chinese ecosystem has delivered a complete AI stack. Like it or not, that's a big news. But what's there not to like when…
I sometimes wonder if there are any security risks with using Chinese LLMs. Is there?
Re: DeepSeek v4
#600It's interesting that they mentioned in the release notes: "Limited by the capacity of high-end computational resources, the current throughput of the Pro model remains constrained. We expect its pricing to decrease significantly once the Ascend 950 has been deployed into production." https://api-docs.deepseek.com/zh-cn/news/news260424#api-%E8%...