Stop giving Anthropic money and figure out how to take the same money to buy some GPUs, and physically insert them into workstations. It is not that hard, I promise.
Uber's $1,500/month AI limit is a useful signal for AI tool pricing
801–810 of 819 posts
Re: Uber's $1,500/month AI limit is a useful signal for AI tool pricing
#802Re: Uber's $1,500/month AI limit is a useful signal for AI tool pricing
#803Earlier quoted context omitted.
Depends, the SMD caps spread across the board the tiny ones do start to fail and go out of spec over time. they are a right pain to replace and hard to spot one that has gone out of spec to cause the chip to start crashing.
Can you not just move the epxensive part (the gpu itself) to a new carrier board in that situation? Also isn't most of the cost of the GPU itself the design of the board, not actually making one, esp if you can move the heat sinks around?
Re: Uber's $1,500/month AI limit is a useful signal for AI tool pricing
#804Re: Uber's $1,500/month AI limit is a useful signal for AI tool pricing
#805Earlier quoted context omitted.
Relative to the current usage demand for tokens is effectively unlimited. If the price of tokens go down people will send more tokens to compensate. We are very very far away from a cost per token where people run out of things they want to send through an LLM.
You’re describing the issue. The problem isn’t that datacenters are under utilized, it’s that they are used to generate something that has a value going down over time, while being financed through debt with the assumption their revenue grow over time. That’s where you have a mismatch. To simplify let’s assume a given datacenter can generate a fixed number of token per second (in reality that depends on the model tok…
Re: Uber's $1,500/month AI limit is a useful signal for AI tool pricing
#806Re: Uber's $1,500/month AI limit is a useful signal for AI tool pricing
#807Earlier quoted context omitted.
Two of the things you’ve listed are some of the most profitable activities in our economy.
I mean, that says a lot about the kind of crisis out current economy is in. How much longer can the United States Be a world leader when it’s primary function is social media and advertising
And writing code is one of the most economically productive activities you can do. Why is it controversial that a technology is good at this?
Re: Uber's $1,500/month AI limit is a useful signal for AI tool pricing
#808Earlier quoted context omitted.
The issue is he’s not actually balanced at all. I’ve never seen him say anything negative about an AI product.
Here's my AI misuse tag: https://simonwillison.net/tags/ai-misuse/ - 54 posts My ongoing coverage of AI ethical issues: https://simonwillison.net/tags/ai-ethics/ - 308 posts I've been the loudest voice about the fundamental insecurity of LLMs for several years: https://simonwillison.net/tags/prompt-injection/ - 150 posts In https://simonwillison.net/2025/Aug/25/agentic-browser-securi... I said "I strongly expect that…
Re: Uber's $1,500/month AI limit is a useful signal for AI tool pricing
#809Earlier quoted context omitted.
Literally in the middle of ripping apart a vibe coded mess at work to figure out what's even worth keeping. Not fun :(
use ai to do that
Re: Uber's $1,500/month AI limit is a useful signal for AI tool pricing
#810Earlier quoted context omitted.
The Mac is very feeble compared to the big iron that the providers run so will be much lower performance. Also many companies would prefer engineers work on the domain problems instead of working on novel LLMs.
I meant “roll your own” LLM for use not build new ones.
Especially concerning the Mac Studio, the GPU is far too weak for enterprise-scale context prefill. You'd need 2 or 4 Studios to process 250k contexts quickly, and even then you'd get bottlenecked by the relatively slow memory bandwidth during the decode stage. It is simply terrible hardware for quick or power efficient inference.