I am not a proper developer and only use AI for faster research of topics so please forgive my ignorance. Could one not save a lot of money on tokens by using the 80/20 or 90/10 rule in that 90% of AI usage is on local models and save that last 10% or less for the frontier models where the local model did not meet the needs? Did they cover this and I misunderstood?
I burned all my tokens researching how to save tokens
31–40 of 237 posts
Re: I burned all my tokens researching how to save tokens
#32Makes sense. LLMs only have 'knowledge' that was encoded by scanning as much data as possible. (Same way a search engine only can find data that was indexed) American LLMs are by and large closed. Chinese are open to a point. And research how these things work is still a big mystery. So yeah, was comprehensive data about saving tokens scanned and indexed? Likely no. So engaging about token saving is going to generate…
Re: I burned all my tokens researching how to save tokens
#33Earlier quoted context omitted.
Trust me there are plenty of us using cloud AI to actually ship stuff. We just aren't writing blog posts about it.
what did you ship?
Re: I burned all my tokens researching how to save tokens
#34Re: I burned all my tokens researching how to save tokens
#35Earlier quoted context omitted.
what did you ship?
Whenever I've answered this question, the reply was always "this is shit", so I don't bother now. Bad faith questions just shouldn't be answered.
Re: I burned all my tokens researching how to save tokens
#36Earlier quoted context omitted.
Trust me there are plenty of us using cloud AI to actually ship stuff. We just aren't writing blog posts about it.
what did you ship?
"Shipped" is a bit strong in my case, but I think this counts as something that isn't mere slop.
Re: I burned all my tokens researching how to save tokens
#37Earlier quoted context omitted.
Whenever I've answered this question, the reply was always "this is shit", so I don't bother now. Bad faith questions just shouldn't be answered.
This just feels like a bad faith comment designed to spike the conversation entirely.
Re: I burned all my tokens researching how to save tokens
#38Claude has been helping me tune models to run better on our big compute rack at work, researching which leading edge open models will fit in hardware and are good for our workloads, and it's been helping me write test fixtures to evaluate how things perform.
At several points, Claude has expressed surprise at the quality of results from my local LLMs. While I know that doesn't mean anything, it still feels like giving Anthropic the finger, which is always great.
Re: I burned all my tokens researching how to save tokens
#39Earlier quoted context omitted.
what did you ship?
Whenever I've answered this question, the reply was always "this is shit", so I don't bother now. Bad faith questions just shouldn't be answered.
doesnt matter what subjective opinions are.
shipped is not pushing something on github.
Re: I burned all my tokens researching how to save tokens
#40I am not a proper developer and only use AI for faster research of topics so please forgive my ignorance. Could one not save a lot of money on tokens by using the 80/20 or 90/10 rule in that 90% of AI usage is on local models and save that last 10% or less for the frontier models where the local model did not meet the needs? Did they cover this and I misunderstood?
Spending thousands of dollars on hardware to save dollars per month on tokens does not make financial sense.
If you run the numbers you'll probably find that using cheaper cloud models makes more financial sense than running those same models locally.