Earlier quoted context omitted.
So what happens when some big bucks VC backed closed source LLM company buys all your compute inventory for the next 5 years? This is not that unlikely. Lambda Labs a little while back was completely sold out of all compute inventory.
I assume it's up to them to say no. They did say they're not in it to make bookoo bucks
Show HN: San Francisco Compute – 512 H100s at <$2/hr for research and startups
161–170 of 189 posts
Re: Show HN: San Francisco Compute – 512 H100s at <$2/hr for research and startups
#162Earlier quoted context omitted.
> Your project has a youthful optimism that I hope you won’t lose as you go. And in fact it might be the way to win in the long run. This is the nicest thing anyone has said to us about this. We're gonna frame this and hang it out on our wall. > So whenever someone comes knocking, begging for a tiny slice of your H100s for their harebrained idea, I hope you’ll humor them. Absolutely! :D
Optimism is (almost) always required in order to accomplish anything of significance. Those who lose it, aren't living up to their potential. I'm not encouraging the false belief that everything you do will work out. Instead I'm encouraging the realization that the greatest accomplishments almost always feel like long shots, and require significant amounts of optimism. Fear and pessimism, while helpful in appropriate…
Re: Show HN: San Francisco Compute – 512 H100s at <$2/hr for research and startups
#163Earlier quoted context omitted.
Couple things, mostly pricing and availability: 1) Margins. Public cloud investors expect a certain margin profile. They can’t compete with Lambda/Fluidstack’s margins. 2) To an extent also big clouds have worse networking for LLM training. I believe only Azure has infiniband. Oracle is 3200 Gbps but not infiniband, same for AWS I believe. GCP not sure but their A100 networking speeds were only 100 Gbps I believe rat…
Low margins and “will this thing still be around in 2 years” are negatively correlated. Where’s the capital for upgrades, repairs, and replacements coming from?
Of course it doesn't always work, and it may be even harder to make it work in the current macroeconomic environment, but it's still pretty standard play.
Re: Show HN: San Francisco Compute – 512 H100s at <$2/hr for research and startups
#164Earlier quoted context omitted.
> Your project has a youthful optimism that I hope you won’t lose as you go. And in fact it might be the way to win in the long run. This is the nicest thing anyone has said to us about this. We're gonna frame this and hang it out on our wall. > So whenever someone comes knocking, begging for a tiny slice of your H100s for their harebrained idea, I hope you’ll humor them. Absolutely! :D
Optimism is (almost) always required in order to accomplish anything of significance. Those who lose it, aren't living up to their potential. I'm not encouraging the false belief that everything you do will work out. Instead I'm encouraging the realization that the greatest accomplishments almost always feel like long shots, and require significant amounts of optimism. Fear and pessimism, while helpful in appropriate…
Re: Show HN: San Francisco Compute – 512 H100s at <$2/hr for research and startups
#165Hi, SF lover [1] here. Anything interesting to note about your name? Will your hardware actually be based in SF? Any plans to start meetups or bring customers together for socializing or anything like that? [1] We have not gone the way of the Xerces blue [2] yet... we still exist! [2] https://en.wikipedia.org/wiki/Xerces_blue
Ah the hardware isn’t gonna be in SF (not the cheapest datacenter space) But I do think a lot of our customers will be out here —- SF is still probably the best place to do startups. We just have so many more people doing hard technical stuff here. Literally every single place I’ve lived in SF there’s been another startup living upstairs or downstairs Good idea to host some in person events!
now that's a hot take if I ever saw one
Re: Show HN: San Francisco Compute – 512 H100s at <$2/hr for research and startups
#166Honest question I don’t know how to consider: are we further along or behind with AI given crypto’s use of GPUs? Has the same cards bought for mining furthered AI, or maybe that demand lead to more research into GPUs and what they can do - or would we be further along if we weren’t wasting these cards on mining?
Ethereum's (thrice delayed) move to PoS put a glut of GPUs on the market, just in time for the AI boom to swallow them back up, so I think it ended up okay. NVDA certainly had a great few days in the market thanks to AI though.
Re: Show HN: San Francisco Compute – 512 H100s at <$2/hr for research and startups
#167Earlier quoted context omitted.
Edge TPUs are low cost, low power inference devices the size of a dime. I have a hundred of them sitting in a closet. (Alas. Anyone want to buy 100 coral minis? :-) The TPUs you rent that are being discussed here are capable of training, consume hundreds of watts and have a heatsink bigger than your fist and really spectacular network links. They're analogous to Nvidia's highest end GPUs from a "what can you do with…
Can I hook a microphone up to a Coral Mini and run Whisper? I'd love to have a home assistant that wasn't on the cloud. As for the rest of them, list them on Amazon and let them do the fulfillment. That $10k of hardware isn't going to sell itself from your closet. (Yet. LLMs are making great strides.)
And that's a good idea, thanks. I've been dreading the idea of using ebay.
Re: Show HN: San Francisco Compute – 512 H100s at <$2/hr for research and startups
#168Earlier quoted context omitted.
Couple things, mostly pricing and availability: 1) Margins. Public cloud investors expect a certain margin profile. They can’t compete with Lambda/Fluidstack’s margins. 2) To an extent also big clouds have worse networking for LLM training. I believe only Azure has infiniband. Oracle is 3200 Gbps but not infiniband, same for AWS I believe. GCP not sure but their A100 networking speeds were only 100 Gbps I believe rat…
What is your differentiator from Lambda? That you are smaller and in a single DC? Sincere question.
Lambda has "Lambda Sprint" which is kinda similar,[1] but Sprint is $4.85/GPU/hr instead of So if you want 128 GPUs for a week, you can't use Lambda reserved (3 year term), you can't use Lambda on-demand (can't get 128 A/H100s on-demand), your options are Lambda Sprint or SF Compute, and SF Compute is offering significantly lower prices.
Re: Show HN: San Francisco Compute – 512 H100s at <$2/hr for research and startups
#169Re: Show HN: San Francisco Compute – 512 H100s at <$2/hr for research and startups
#170Earlier quoted context omitted.
I assume it's up to them to say no. They did say they're not in it to make bookoo bucks
I must say, this is the worst I've seen "beaucoup" spelled.