Live data from Hacker News

Higher usage limits for Claude and a compute deal with SpaceX

anthropic.com

221–230 of 519 posts

Re: Higher usage limits for Claude and a compute deal with SpaceX

#221

Earlier quoted context omitted.

[flagged]

It may be more productive to ask what is right with burning fossil fuels for electricity right in the middle of marginalized communities that have to bear the cost of this pollution for AI slop.

Marginalised communities.

Righteo, I guess I better suspend my white privilege.

Re: Higher usage limits for Claude and a compute deal with SpaceX

#222

"use all of the compute capacity at their Colossus 1 data center" So, they handed out all of their data center to Anthropic; Grok wasn't using it much?

Moved to Colossus 2. Though I guess you could still frame it as 'don't they need Collosus 1 AND 2' if you want...

Re: Higher usage limits for Claude and a compute deal with SpaceX

#223

> 300 megawatts of new capacity (over 220,000 NVIDIA GPUs) The scale is just mindboggling here. Are there any blog posts or anything discussing what kind of infrastructure is used for even just the inference side (nevermind the training) for SotA models like Opus? I would have thought it might be secret, but given that you can actually run the models yourself on AWS Bedrock doesn't that give an indication?

How many instances of Doom can it run though?

Re: Higher usage limits for Claude and a compute deal with SpaceX

#224

Earlier quoted context omitted.

Pretty smart for SpaceX though. They’re turning an asset they made for a money-pit (Grok) into probably a major source of revenue ahead of their IPO.

We all remember 2 weeks ago when SpaceX bought $10B of Cursor services. https://news.ycombinator.com/item?id=47855293 Since Cursor often relies on Claude models, some of those services will flow back to their own datacenter compute. Especially if there's, lets call it, "customer demand loadbalancing optimization agreements" that makes those Cursor services prioritize Claude models using the app keys that get load-bal…

Either way, now those datacenters run Claude that they didn't before.

Re: Higher usage limits for Claude and a compute deal with SpaceX

#225

Earlier quoted context omitted.

Pretty smart for SpaceX though. They’re turning an asset they made for a money-pit (Grok) into probably a major source of revenue ahead of their IPO.

I see it more of lets make money off the hardware we are not using anymore. From Elon on X: ... After that, I was ok leasing Colossus 1 to Anthropic, as SpaceXAI had already moved training to Colossus 2. https://x.com/elonmusk/status/2052069691372478511

But like... most companies are so short of GPUs they'd run it on anything. SpaceXAI not needing the compute is not really a good sign imo.

Re: Higher usage limits for Claude and a compute deal with SpaceX

#226
post #176

Earlier quoted context omitted.

We all remember 2 weeks ago when SpaceX bought $10B of Cursor services. https://news.ycombinator.com/item?id=47855293 Since Cursor often relies on Claude models, some of those services will flow back to their own datacenter compute. Especially if there's, lets call it, "customer demand loadbalancing optimization agreements" that makes those Cursor services prioritize Claude models using the app keys that get load-bal…

I don't think it's the conspiracy theory that you're making it out to be. It is publicly known that the vast majority of deals in the AI space are circular in nature without the need for explicitly encoding any of it in a legal contract or even tacit agreements. e.g. Nvidia has invested significantly in many AI companies including both Anthropic and OpenAI which rely heavily on Nvidia's hardware and will undoubtedly…

Your argument is that since it is common in a bubble to make circular deals, there is no conspiracy. But you seem to suggest that people committing tens of billions of dollars aren’t looking any further down the pipeline than the name on the receiving bank account? Have you ever been anywhere near a large deal?

Re: Higher usage limits for Claude and a compute deal with SpaceX

#227
post #49

Doubling the five-hour rate limits is merely a marketing stunt if the weekly rates are not also doubled. It simply means that you can reach the weekly limits in three days instead of five.

I don’t think I’ve hit either limit a single time in the past 5 months after upgrading to the $100 plan. On heavy weeks I probably am using it consistently for at least 6+ hours a day. Although, I’m pretty rigorous about always keeping my sessions under 200-250k tokens.

I've maxed out weekly limits for 2 $200 accounts before

Re: Higher usage limits for Claude and a compute deal with SpaceX

#229

> 300 megawatts of new capacity (over 220,000 NVIDIA GPUs) The scale is just mindboggling here. Are there any blog posts or anything discussing what kind of infrastructure is used for even just the inference side (nevermind the training) for SotA models like Opus? I would have thought it might be secret, but given that you can actually run the models yourself on AWS Bedrock doesn't that give an indication?

I know you're probably talking about the compute infrastructure, but I think the electricity infrastructure side is interesting too, data centers are doing things in dumb ways because the need for operational expansion speed is greater than the need dollars:

> It’s regulation with the utilities. There are ramp rates, there are all of these things that you’re supposed to do to not screw up the grid. Data centers have been in gross violation of that. When you think about what’s wrong with data centers, they have load volatility, which we just talked about, then they decide to power it with behind-the-meter natural gas generators. These natural gas generators, their shaft is supposed to last for seven years. It’s lasting 10 months because of all the cycling.

https://www.volts.wtf/p/doing-data-centers-the-not-dumb-way

On the compute infrastructure, there are standard NVIDIA reference designs like this:

https://www.nvidia.com/en-us/technologies/enterprise-referen...

I haven't bothered to look but I'd guess Mellanox GPU-to-GPU networks, and massive custom code for splitting tensors across GPUs, and for shuttling activations across GPU nodes.

Re: Higher usage limits for Claude and a compute deal with SpaceX

#230
post #192

Earlier quoted context omitted.

As I understand it, the problem is cooling. There isn't any medium to take away the heat, so the only option is to slowly radiate it away.

Anyone who has googled just once to ask if datacenters in space make any sense, has found out they don't because they can't get rid of heat. That leaves only two kinds of people left who are still talking excitedly about datacenters in space: The uninformed and the grifters.

The existence of starlink proves that this is false. Look at most current pitches, they don’t talk about GW-class monsters anymore. There’s absolutely nothing stopping a 20-30kW satellite bus the size of starlink (or I guess up to 100kW? once starship is available - it’s all about payload fairing diameter) from hosting ~1 rack of compute and antennas. The economics may or may not make sense, we’ll have to see.

There’s very little research work needed to make this happen; it’s all about engineering some satellite buses and having them fly in close formation to get a “data center”. And this group of satellites in sun-synchronous orbit would relay to a comms constellation e.g. starlink itself) and operate as a global scale data center. The heat management and orbital mechanics are all straight forward really.

Post reply on HN