Live data from Hacker News

Outsourcing plus local AI will soon become more economical vs. frontier labs

signalbloom.ai

171–180 of 408 posts

Re: Outsourcing plus local AI will soon become more economical vs. frontier labs

#171

Earlier quoted context omitted.

[flagged]

People can simultaeneously be reprehensible idiots while being a reliable expert on something they have personally invested billions of dollars into and operate at scale.

I'm not saying he actually is an expert, but he could be an expert and still lie for any number of reasons.

Re: Outsourcing plus local AI will soon become more economical vs. frontier labs

#172

Earlier quoted context omitted.

[flagged]

People can simultaeneously be reprehensible idiots while being a reliable expert on something they have personally invested billions of dollars into and operate at scale.

Elon is a specialist of lying about stuff he invested billions in to make it look more valuable than it is (he's been doing that for Tesla for years). It's not a lack of expertise, it's the lack of any sense of integrity (and self respect).

He's lagging the AI race despite having tons of compute available, so he tries to make a narrative about how it's not that the model is behind, it's just smaller than the competition.

Re: Outsourcing plus local AI will soon become more economical vs. frontier labs

#173
post #93
post #48

When discussing LLM pricing, people are missing the plot. The subscription token price is 10x-40x cheaper than API pricing. Your 90$ Claude subscriptions give you close to $1000 to $4000 in equivalent API token pricing. The second issue is that the quality of the model “operator” makes a massive difference in the outcomes. Highly skilled senior devs who know how to prompt and have high agency will outperform team peo…

> Lastly, there is a massive difference in capabilities, determinism, and error handling between 5T SOTA models like Opus What's your source for Opus being a 5T model? > and tiny distillations from DeepSeek that perform well only in benchmarks. I don't think you know what you're talking about. Local models aren't “distillations from Deepseek”. And they don't perform well “only in benchmarks”, Qwen 3.6 is a very decen…

https://arxiv.org/abs/2604.24827

From this paper

Re: Outsourcing plus local AI will soon become more economical vs. frontier labs

#174

Earlier quoted context omitted.

Also, your local hardware is in no way capable of running the types of models that the cloud providers do, it’s just not economically feasible, and it never will be.

Depends on what you mean by "economically feasible". Even very cheap mini-PCs and laptops can run any of the models run by cloud providers, albeit at a much lower speed (i.e. with the weights stored on SSDs). Whether such a low speed is useful, depends on the application. For something like a coding assistant or bug scanning, an instant response is desirable, but certainly not necessary.

The SSD would wear out in days while the laptop generates two responses a day. This is like saying you could power your home with AA batteries, yes technically you could but in practice entirely infeasible.

Re: Outsourcing plus local AI will soon become more economical vs. frontier labs

#175
post #61

Earlier quoted context omitted.

Yes and no. Just take a look at the OpenRouter providers page: https://openrouter.ai/deepseek/deepseek-v4-pro/providers Deepseek v4 Pro is much cheaper when provided by Deepseek itself, likely as a combination of the loss leader strategy you mention and the desire to have more data flow through their pipeline for training. However, the same open weights model, provided by other providers, is somewhere in the $2-3/1M…

> Unless you mean that releasing open weights models is the loss leader, in which case, you might be right but I hope you're wrong. This is specifically what I meant. DeepSeek’s official service is trying to recoup some of the training and engineering costs too. The other providers only have to recoup their hardware costs and the cost of a team to run it. Even though DeepSeek’s official service is more expensive per…

Except DeepSeek's official service is _less_ expensive per token, which suggests they're underpricing it substantially as well to attempt to draw more attention / more data.

Re: Outsourcing plus local AI will soon become more economical vs. frontier labs

#176

Earlier quoted context omitted.

Depends on what you mean by "economically feasible". Even very cheap mini-PCs and laptops can run any of the models run by cloud providers, albeit at a much lower speed (i.e. with the weights stored on SSDs). Whether such a low speed is useful, depends on the application. For something like a coding assistant or bug scanning, an instant response is desirable, but certainly not necessary.

The SSD would wear out in days while the laptop generates two responses a day. This is like saying you could power your home with AA batteries, yes technically you could but in practice entirely infeasible.

Weights are write-once data.

Re: Outsourcing plus local AI will soon become more economical vs. frontier labs

#177

Great article that reinforces my own opinion but adding the cleverness of adding low cost human labor into the equation. Nice. I spent a month comparing Gemini Ultra plan to using much lower cost DeepSeek v4 with open source coding harnesses and, spoiler alert: I was happier using the much cheaper and more environmentally friendly open models: https://marklwatson.substack.com/p/my-evaluation-of-ai-agent...

FYI, Gemini is the most environmentally friendly model.

https://share.google/aimode/a0O95wzk2UUhIXLUI

https://cloud.google.com/blog/products/infrastructure/measur...

Re: Outsourcing plus local AI will soon become more economical vs. frontier labs

#178

Earlier quoted context omitted.

> The quality of the model “operator” makes a massive difference in the outcomes. My hunch is that this is the source of much of the variability in outcomes upstream of HN commenters claiming extremes of, "This model changes everything!" to "This[same] model is crap." We haven't operationalized what it means to "be good at prompting," nor developed proxies/heuristics/shibboleths for accessing prompting skill. There's…

It's 100% this. Many people suck at prompting. It's likely that habits from search are ingrained. But in general some people are just so bad at it .

Prompting is just writing specification documents. A lot of people are very bad at this. I suppose that more to the point, a lot of people are just bad at writing.

Re: Outsourcing plus local AI will soon become more economical vs. frontier labs

#179

I have really been trying to get local models to work. I have tried different harnesses, tooling, skills, prompts, etc. But when I compare claude code with anthropic models or codex with gpt 5.5, vs qwen, glm or gemma and the same harnesses, the frontier models come out massively ahead. I am at the point where I just don't see the point of the non-frontier models, they waste more time than they save.

I came to the same conclusion. For the amount that a query costs, using Opus all the time is the cheapest option.

Re: Outsourcing plus local AI will soon become more economical vs. frontier labs

#180

Earlier quoted context omitted.

its a private company, what exactly do you expect to 'see'?

Anthropic IPO's in less than 5 months and I guarantee you any company that officially is in the black will proudly shout it from the rooftops.

> Anthropic IPO's in less than 5 months

pure speculation. about as valuable as my linked wsj reporting i suppose. given thats the case, maybe you shouldnt claim so confidently that they are money incinerators.

Post reply on HN