Live data from Hacker News

Show HN: sllm – Split a GPU node with other developers, unlimited tokens

sllm.cloud

91–100 of 117 posts

Re: Show HN: sllm – Split a GPU node with other developers, unlimited tokens

#92

Earlier quoted context omitted.

Isn't this a bad deal? Or is there an error in my math? For $40, I'd get 20 tok/s * 2.6M seconds per month = 52M tokens of DeepSeek v3.2 per month if I run it 24/7, which is not realistic for most workloads. On OpenRouter [1], $40 buys 105M tokens from the same model, which is more than 52M tokens, and I can freely choose when to use them. [1]: https://openrouter.ai/deepseek/deepseek-v3.2

20 tok/s is an average. It can be more, it can be less. If you are running off-peak I'm sure you'd get some crazy number.

That doesn’t matter when you have the average. Even if you are somehow able to get 10000tok/s during off peak times, by virtue of how averages work, you’re still only getting 52M tokens per month (as calculated above).

Re: Show HN: sllm – Split a GPU node with other developers, unlimited tokens

#93
post #84

Earlier quoted context omitted.

> $40/mo for deepseek r1 seems steep compared to a pro sub on open ai /claude unless you run 24x7. "Running 24x7" is what people want to do with openclaw.

Seems like they have a rate limit so it is kinda the same as normal subs - don’t really see the advantage yet

Presumably the rate limit is much higher

Re: Show HN: sllm – Split a GPU node with other developers, unlimited tokens

#94
post #84

Earlier quoted context omitted.

> $40/mo for deepseek r1 seems steep compared to a pro sub on open ai /claude unless you run 24x7. "Running 24x7" is what people want to do with openclaw.

Seems like they have a rate limit so it is kinda the same as normal subs - don’t really see the advantage yet

> Seems like they have a rate limit so it is kinda the same as normal subs - don’t really see the advantage yet

It's not really the same "limit", AIUI.

SLLM: Being capped to the rate would make your openclaw run slowly, but still be able to work 24x7.

Normal Subs: Hiting the limit means your openclaw doesn't run at all for hours.

Re: Show HN: sllm – Split a GPU node with other developers, unlimited tokens

#96

This is great, thanks! I personally would like something like this but with "regular" GPU access. Some people still use them for something other than LLMs ^^.

hotaisle.xyz has amd mi300x VMs for $1.99/gpu/hr. on-demand, billed by the minute.

(i'm the ceo)

Re: Show HN: sllm – Split a GPU node with other developers, unlimited tokens

#98

1 week or even 1 day windows would be great, especially just to test it at this early stage

There are a number of on-demand providers of GPU compute out there. It is relatively straightforward to run inference on them.

I've got a box of 8x MI300x sitting around waiting for stuff like this.

$128...

8x MI300X Bare Metal - $15.94/hour (1 available) CPU: Xeon Platinum 8462Y+ (64 cores) • Memory: 2.0 TiB • Disk: 124 TB • Minimum Reservation: 8 hours

ssh admin.hotaisle.app

(i'm ceo)

Re: Show HN: sllm – Split a GPU node with other developers, unlimited tokens

#100
post #90

[flagged]

For my part, the code quality of the next.js dashboard isn't even something I'd evaluate. I instantly get a quick, functional-appearing view of the offering. I can picture how I might interact with it and what mental gymnastics stand between right now and me pulling out a credit card. I don't see marketing fluff that bring up more questions than it provides answers. I can be pretty certain I won't wind up in a sales…

Right but without at least the effort of a sales funnel I have zero confidence that this guy has thought anything through.

This is a single afternoon vibe-coded project by all indications.

Post reply on HN