Live data from Hacker News

Deepseek R1-0528

huggingface.co

41–50 of 264 posts

Re: Deepseek R1-0528

#41
post #12

Earlier quoted context omitted.

On the day Nvidia report earnings too. Pretty sure it's just a coincidence, bro.

Yeah the timing seems strange. Considering how much money will move hands based on those results this might be some kind of play to manipulate the market at least a bit.

Plenty of manipulation to go around..

"Tech Chip software stocks sink on report Trump ordered halt to China sales" - https://www.cnbc.com/2025/05/28/chip-software-trump-china.ht...

Re: Deepseek R1-0528

#42

Out of sheer curiosity: What’s required for the average Joe to use this, even at a glacial pace, in terms of hardware? Or is it even possible without using smart person magic to append enchanted numbers and make it smaller for us masses?

You can run the 4bit quantized version of it on a M3 Ultra 512GB. That's quite expensive though. Another alternative is a fast CPU with 500GB of DDR5 RAM. That of course, is also not cheap and slower than the M3 Ultra. Or, you buy multiple Nvidia cards to reach ~500GB of VRam. That is probably the most expensive option but also the fastest

If you use the excess memory for AI only it's cheaper to rent . A single H100 costs less than $2 per hour. (incl power)

Re: Deepseek R1-0528

#43

Earlier quoted context omitted.

Not every prompt is privacy sensitive. For example you could use it to summarize a public article.

Every prompt is valuable.

And you are getting something valuable in return. It's probably a good trade for many, especially when they are doing something like summarizing a public article.

Re: Deepseek R1-0528

#46
post #4

I love how Deepseek just casually drops new updates (that deliver big improvements) without fanfare.

Honest question, how do you know this is a big improvement? Are there any benchmarks anywhere?

There will be a video from FireShip if its a big one. /s

Re: Deepseek R1-0528

#47
post #42

Earlier quoted context omitted.

You can run the 4bit quantized version of it on a M3 Ultra 512GB. That's quite expensive though. Another alternative is a fast CPU with 500GB of DDR5 RAM. That of course, is also not cheap and slower than the M3 Ultra. Or, you buy multiple Nvidia cards to reach ~500GB of VRam. That is probably the most expensive option but also the fastest

If you use the excess memory for AI only it's cheaper to rent . A single H100 costs less than $2 per hour. (incl power)

Vast.ai has a bunch of 1x H100 SXM available, right now the cheapest at $1.554/hr.

Not affiliated, just a (mostly) happy user, although don't trust the bandwidth numbers, lots of variance (not surprising though, it is a user-to-user marketplace).

Re: Deepseek R1-0528

#48
post #10

No information to be found about it. Hopefully we get benchmarks soon. Reminds me of the days when Mistral would just tweet a torrent magnet link

i think usually deepseek posts a paper after a model release about a day later.

no idea why they cant just wait a bit to coordinate stuff. bit messy in the news cycle.

Re: Deepseek R1-0528

#49
post #10

No information to be found about it. Hopefully we get benchmarks soon. Reminds me of the days when Mistral would just tweet a torrent magnet link

Benchmarks seem like a fools errand at this point; overly tuning models just to specific test already published tests, rather than focusing on making them generalize.

Hugging face has a leader board and it seems dominated by models that are finetunings of various common open source models, yet don't seem be broader used:

https://huggingface.co/open-llm-leaderboard

Re: Deepseek R1-0528

#50
post #48
post #10

No information to be found about it. Hopefully we get benchmarks soon. Reminds me of the days when Mistral would just tweet a torrent magnet link

i think usually deepseek posts a paper after a model release about a day later. no idea why they cant just wait a bit to coordinate stuff. bit messy in the news cycle.

honestly a power move.

it's almost as if they don't care about creating a proper buzz.

Post reply on HN