It's amazing that it took months to figure this out. "Well we thought that if engineers are told to maximize costs through AI use, to consume as much as possible of a resource that costs us money, then obviously good things will happen. Imagine my surprise when it didn't turn out that way." Imagine if engineers were ranked based on their AWS spend. People allocate VMs and fill databases with terabytes of random bits,…
Uber’s COO says it’s getting harder to justify money spent on tokenmaxxing
41–50 of 360 posts
Re: Uber’s COO says it’s getting harder to justify money spent on tokenmaxxing
#42It's amazing that it took months to figure this out. "Well we thought that if engineers are told to maximize costs through AI use, to consume as much as possible of a resource that costs us money, then obviously good things will happen. Imagine my surprise when it didn't turn out that way." Imagine if engineers were ranked based on their AWS spend. People allocate VMs and fill databases with terabytes of random bits,…
The point of this was always to explore what is possible with AI as quickly as possible. Obviously, there is going to be a lot of waste, but the 5-10% of employees who are truly thinking about it and discovering novel applications are what you are truly after. Because right now, you effectively have a giant, as of yet poorly explored space of potential uses. Anyone who can find the actually valuable portions of the s…
OTOH maybe we’re in for a future of patenting prompts.
Re: Uber’s COO says it’s getting harder to justify money spent on tokenmaxxing
#43It's amazing that it took months to figure this out. "Well we thought that if engineers are told to maximize costs through AI use, to consume as much as possible of a resource that costs us money, then obviously good things will happen. Imagine my surprise when it didn't turn out that way." Imagine if engineers were ranked based on their AWS spend. People allocate VMs and fill databases with terabytes of random bits,…
You say "amazing that it took months to figure this out" as if the answer to the question is obvious. But it's not. Some FAANGs are doing amazing things with unlimited tokens. Other companies have no clue what to do with tokens, they've just told their engineers to max them. It really depends on how you're using the tokens . If you're just using them for Codex and Claude Code - yeah, tokenmaxxing is incredibly dumb.
Would love to know what things!
Re: Uber’s COO says it’s getting harder to justify money spent on tokenmaxxing
#44I find it useful that if they cut the use altogether I will pay for it out of pocket.
The former is the issue, and how many companies have been operating. It's like a trucking company ranking driver effectiveness by fuel used instead of by cargo moved.
Re: Uber’s COO says it’s getting harder to justify money spent on tokenmaxxing
#45As soon as tokens stop stop being subsidized, heavy agentic use will become as least as expensive than paying an (entry level) employee. When this happens many companies will trade off havy tolen usage for (maybe a bit slower, bit less accurate) employees again.
I also don't think that blitz scaling will work like with Uber. The engineers are still there. We can work without the LLM tools.
Re: Uber’s COO says it’s getting harder to justify money spent on tokenmaxxing
#46Re: Uber’s COO says it’s getting harder to justify money spent on tokenmaxxing
#47Now we are going to get a new profession. Token Engineer! They will be experts on tokenmaxxing! The job growth that the billionaire CEOs promised us from AI is finally here!
Re: Uber’s COO says it’s getting harder to justify money spent on tokenmaxxing
#48Re: Uber’s COO says it’s getting harder to justify money spent on tokenmaxxing
#49Maybe don't use the most expensive models on the planet? Maybe use AI like a tool and not this black box that grants wishes?
Re: Uber’s COO says it’s getting harder to justify money spent on tokenmaxxing
#50What if... we stop for a moment, and then, after thinking for a moment, we stop hammering nails with a microscope, and stop using token usage as a metric of productivity? I know it's sounds stupid, but what if