Live data from Hacker News

Show HN: Echo – Fable-level results at 1/3 the cost using open-weight models

news.ycombinator.com

181–190 of 254 posts

Re: Show HN: Echo – Fable-level results at 1/3 the cost using open-weight models

#181

Grandiose claims that are hardly backed up at all. I find that very off-turning!

Have you taken a look at https://echo.tracerml.ai/eval/ ? It looks promising for sure!

None of those are SWE.

And still Fable beats it hands down 8-0 in one of them, and is at worst even in some others.

Also it doesn't make logical sense: A router can save costs, yes, but not magically be "smarter" somehow.

That's like selling "free energy".

Re: Show HN: Echo – Fable-level results at 1/3 the cost using open-weight models

#183
post #156
post #135

Earlier quoted context omitted.

People point to the equivalent API costs to show that they are getting a great deal on the subscription, 10,000 dollars worth of tokens for 200 dollars. I do wonder if it's the other way around though - are the API users simply getting ripped off? I have seen Dario say in multiple interviews that they are profitable on inference, which maybe he was only meaning to refer to API usage, but that's not the impression I g…

> I'm not convinced that the 200 dollar plans are unprofitable. Especially considering not everyone is tokenmaxxing, and in most parts of the world people take leave and companies do not cut their subscriptions. I suspect they are priced to have a lifetime average price/token amount that is roughly break-even, or maybe a slight loss leader. > have seen Dario say in multiple interviews that they are profitable on infe…

> Especially considering not everyone is tokenmaxxing

Exactly. I have the Claude $100/mo plan, and use it moderately for open source hobby stuff. I still haven't dipped my toes into the Fable pool, but I always use Opus 4.8 on xhigh, and I never hit my limits.

On the other hand, though, there have been times when I've looked at /usage for a long-running session (e.g., 7-10 days, after it's compacted a few times), and it showed I'd used ~$450 worth of tokens just for that session. So I'm clearly getting value for the money here when it comes to the subscription cost. But I still don't hit limits, so...

Re: Show HN: Echo – Fable-level results at 1/3 the cost using open-weight models

#184

Earlier quoted context omitted.

All enterprises users (people using them for work and not side projects) can't get the subsidized plans. I would say subsidized plans are a minority of usage?

People _can_ get subsidized plans for work: we use Claude Teams, $100/mo premium seat, which caps at 150 seats. Not enterprise tier, but fine for SMBs.

As long as you are fine with everything your team puts into Claude being retained by Anthropic. Afaik only enterprise API plans provide non data retention policies.

Re: Show HN: Echo – Fable-level results at 1/3 the cost using open-weight models

#185
So it is a mixture of models, then each model have a mixture of experts, making it mixture of mixture of experts...why don't we just make it simple and have a central expert router made up of random forest...all we need some kind of actor system to share the gradients of the experts...and ensemble it

Re: Show HN: Echo – Fable-level results at 1/3 the cost using open-weight models

#186
post #180

Earlier quoted context omitted.

I believe they will last until they IPO, and not long after that. $200/mo plans are not good for their P&L when their users using $10000 worth api credits. That's -98% margin loss per user.

No way in hell are the majority of Claude Code users burning 10k worth of credits. Many of them probably barely use it. There'll be a bell curve, and we have no idea what it looks like.

They don't need to be the majority. One big company paying 200/300k in credits each month easily makes up for the majority of single users not doing so. I believe AI companies today make money through b2b enterprise deals and not selling to individual users, the latter is mostly a marketing expense to get people to use their product instead than the competitors one.

Re: Show HN: Echo – Fable-level results at 1/3 the cost using open-weight models

#187

No benchmarks, no info on which models are used, ai generated video, just a signup page with nothing else. Anyhow, this kinda reminds me of that quote about architecture: "We replaced our monolith with micro services so that every outage could be more like a murder mystery."

[dead]

Re: Show HN: Echo – Fable-level results at 1/3 the cost using open-weight models

#189

Earlier quoted context omitted.

I am too young (most of us on here are) to have lived through the paying for time on time-share machines in the 60s/70s, but this is giving me creepy memories of paying for sprintnet/telenet and tymnet... And I guess aol, compuserv, delphi. Are we really doing this computing model again?

If I remember well, visual studio/ msdn used to be like 5k-10k per year... Basically any IT tool cost a shit load of money. I have in mind 50k-100k ish for 3d studio max or was it softimage? (Well seems softimage https://www.awn.com/animationworld/siggraph-news-announcing-... ). So... Basically we are back to this era.

I think you were able to buy VS Pro perpetual licenses for one time 500-1000 at all times. Of course they had higher tiers and a subscription model which were more expensive but in reality for most people there was little to no added value.

Re: Show HN: Echo – Fable-level results at 1/3 the cost using open-weight models

#190

I wouldn't be surprised if "the best model" becomes a niche concept over the next few years. For most production systems the winning architecture may end up being an orchestrator that knows when to call a cheap model when to escalate to a stronger one and when to combine multiple outputs.

Isnt that the idea from gemini cli?
Post reply on HN