Live data from Hacker News

I burned all my tokens researching how to save tokens

quesma.com

141–150 of 237 posts

Re: I burned all my tokens researching how to save tokens

#141
post #92

Earlier quoted context omitted.

shipped = paying real users ( or a non profit with users) doesnt matter what subjective opinions are. shipped is not pushing something on github.

My definition would be "users waiting for it who received it"; that's the one that fits the broadest real world definition of the word. Not everything is an organisation, not everyone is paying.

If the only user is the author, then I don’t see a meaningful distinction between shipped and not shipped. It’s like a blacksmith who said they invented a new tool with great value but they built it themselves and no one else is willing to use it because it’s so customized to their workflow.

Edit: I’ll add a caveat that I’d agree it was shipped if it’s making the author money even if they are the only user, but otherwise it all smells like dotcom era Pets.com level of selling a dollar for 99 cents and claiming it’s the future.

Re: I burned all my tokens researching how to save tokens

#142

Earlier quoted context omitted.

Why would anyone pull it? When I want slop I just ask the AI to generate some exactly the way I like it.

I get the skepticism, I really do, but my hope is that your own AI(s) confirm exactly what I'm claiming within these comments. Like I've told others, there's no need to take my word for it... the entire point of the project is that anyone can execute/prove the same things I claim using their own hardware via a Doom-replay type format. If I was one of those influencer types, then I would've simply led with "it's an ag…

> I get the skepticism, I really do, but my hope is that your own AI(s) confirm exactly what I'm claiming within these comments

Given that you stated your readme is for bots, given that you expect our own AI to interpret it based on your evidence, how do I, as a human, go about evaluating this project without just writing it off as slop from the linked evidence?

What is the starting off point for a human, not an AI agent, to evaluate this?

And to be clear if the answer is to just download arbitrary code onto my machine and run it, that’s not good enough given the increase in attacks from AI projects and fake recruiters telling you to just run it bro. I’ve already had to pass on multiple interviews despite being unemployed because the human/ai on the other side demanded root access or for me to run arbitrary code onto my machine as part of their process.

Re: I burned all my tokens researching how to save tokens

#143

The author touches on an issue which bothered me, which was the thought of many agents re-solving the same issues over and over. I see a few comments here too, mentioning fixing issues which may already have a solution elsewhere. That was why I created https://pushrealm.com which started as essentially a Stackoverflow clone via MCP. It has now become a way for agents to converge on complete, shared answers for emergi…

Elsewhere in these replies, someone linked to very similar idea by Karpathy https://gist.github.com/karpathy/442a6bf555914893e9891c11519...

Re: I burned all my tokens researching how to save tokens

#144
post #33

Earlier quoted context omitted.

Whenever I've answered this question, the reply was always "this is shit", so I don't bother now. Bad faith questions just shouldn't be answered.

shipped = paying real users ( or a non profit with users) doesnt matter what subjective opinions are. shipped is not pushing something on github.

I have used AI to ship software that users pay for.

Re: I burned all my tokens researching how to save tokens

#145
post #31
post #23

I am not a proper developer and only use AI for faster research of topics so please forgive my ignorance. Could one not save a lot of money on tokens by using the 80/20 or 90/10 rule in that 90% of AI usage is on local models and save that last 10% or less for the frontier models where the local model did not meet the needs? Did they cover this and I misunderstood?

Unlees you invest thousands your local ai won't even come close to the cheap cloud llms. Local is only worth it if you care about privacy or have a legitimate usage for the hardware otherwise. Money wise it isn't worth it

I mean, if you're using one of the $200/month plans, you'd recoup the cost in a year or two.

Re: I burned all my tokens researching how to save tokens

#146
post #23

I am not a proper developer and only use AI for faster research of topics so please forgive my ignorance. Could one not save a lot of money on tokens by using the 80/20 or 90/10 rule in that 90% of AI usage is on local models and save that last 10% or less for the frontier models where the local model did not meet the needs? Did they cover this and I misunderstood?

This is like the "half of my marketing spend is wasted" quote. The complexity is finding out which half.

We could probably get the same results by offloading 20% of obviously smaller / simpler / well defined tasks to smaller models and keep the 80% for ones that benefit from the big iron. Not the 80/20 they were referring to, but still it's something.

I find myself wasting time on smaller models or wasting money on frontier models. I only have so much of either.

Re: I burned all my tokens researching how to save tokens

#147
post #29

Earlier quoted context omitted.

This is like the "half of my marketing spend is wasted" quote. The complexity is finding out which half.

I think the idea or methodology is something along the line of one starts with the local models then when they hit a wall then continue with their current results in a frontier model, thus the other half is those last bits one could not compute locally.

If you are only using chatbots for code gen then for sure this is a reasonable approach. If you're trying to learn something then your best bet is to not have the lowest jpg quality version of human knowledge be your teacher.

Re: I burned all my tokens researching how to save tokens

#149
post #145
post #31

Earlier quoted context omitted.

Unlees you invest thousands your local ai won't even come close to the cheap cloud llms. Local is only worth it if you care about privacy or have a legitimate usage for the hardware otherwise. Money wise it isn't worth it

I mean, if you're using one of the $200/month plans, you'd recoup the cost in a year or two.

You're not running an Opus 4.8/GPT 5.6 tier model on $5k of hardware at a useful Tok/s

Re: I burned all my tokens researching how to save tokens

#150

Earlier quoted context omitted.

Trust me there are plenty of us using cloud AI to actually ship stuff. We just aren't writing blog posts about it.

what did you ship?

I shipped this shader based video compositor as part of a bigger project ive been working on since before LLMs but the compositor was 100% ai and it works fantastically. It's not a huge project but I and a number of my users use it. https://lowkeyviewer.com/studio/
Post reply on HN