Live data from Hacker News

LLMs: Intelligence vs. Cost

openteams.com

41–48 of 48 posts

Re: LLMs: Intelligence vs. Cost

#41
post #10

Complaining about "bad charting" and posting a chart with y-axis that doesn't start at 0 is kinda weird.

Starting at 0 is really only useful if you’re plotting a ratio measure and not just an interval one (https://en.wikipedia.org/wiki/Level_of_measurement#Interval_... ), which I’m not sure this “intelligence index” is.

Re: LLMs: Intelligence vs. Cost

#42
post #39
post #38

Earlier quoted context omitted.

Eh no. Expecting a non-programmer to be able to generate this is very elitist. A static image is still the standard.

The author's job description is literally "Staff Software Engineer at OpenTeams. Dask maintainer."!

The discussion isn't about what one person should do, but on what is appropriate in general for depicting such graphs. What he did do is in line with current and long standing recommendations.

Re: LLMs: Intelligence vs. Cost

#43

Earlier quoted context omitted.

I think the key there is “hardware you already own.” If you own a graphics card you bought for gaming or a laptop you bought for doing schoolwork there is $0 in cost of local AI tokens, because 100% of the cost was assigned to doing other things.

It's bad accounting to only look at marginal cost and ignore depreciating hardware asset costs.

[deleted]

Re: LLMs: Intelligence vs. Cost

#45
post #8

This looks great! I also think speed should be part of the metric (i.e. how long does the model take to actually solve a task). For me, I prefer to run expensive models such as Sol on light reasoning, which usually gives me good answers with quick responses. For my style of coding (quick back-and-forths and corrections) it makes a big difference if a model comes back in 1-2 minutes compared to 5-10, and I am happy to…

Artificial analysis, the place where they get this data from, has a cost vs time chart.

Go to https://artificialanalysis.ai/ and scroll down to the second graph under “Speed & Latency”.

I think this is the most import graph on their page. I wish they would let us filter by intelligence, or pass rate, and then see this graph. This is the tradeoff that actually matters, cost/token or tok/s can be very misleading (take glm-5.3-flash as an example).

Re: LLMs: Intelligence vs. Cost

#46

The complaint about not being able to switch between linear and log is valid, which is what I did for making a 3D speed/cost/quality frontier application for a recent meetup talk: https://www.williamangel.net/apps/model_performance.html Because speed is important, as the reasoning and hardware determine both cost and speed. it's a three dimensional tradeoff.

I really like the 3D version, but I strongly believe you need to consider the number of tokens required to complete a task, it heavily impacts the results for certain models that rely heavily on test time compute (Glm-5.3-flash is the newest example).

Re: LLMs: Intelligence vs. Cost

#47

Earlier quoted context omitted.

I think the key there is “hardware you already own.” If you own a graphics card you bought for gaming or a laptop you bought for doing schoolwork there is $0 in cost of local AI tokens, because 100% of the cost was assigned to doing other things.

It's bad accounting to only look at marginal cost and ignore depreciating hardware asset costs.

[deleted]

Re: LLMs: Intelligence vs. Cost

#48
post #2

Open Teams originally wanted to rent out open source developers to sponsors with Oliphant controlling everything. Now they pivoted to installing local LLMs (on what hardware exactly?). What will happen is that this will be the third consultancy with a lofty narrative after Enthought and Anaconda that Oliphant established. It is always bait-and-switch.

Travis Oliphant here. That is very negative and completely unfounded take on my efforts to hire, fund, and redirect resources to open-source developers. I'm saddened and disheartened that you would feel that way. I didn't control Enthought. I founded but never controlled Anaconda on my own and my influence there has been small since 2018 and nearly non-existent since 2021. OpenTeams is another joint effort with multiple stake holders working to help mission-driven engineers work with investors to bring distributed and owned intelligence to everyone. I'm driven to create more owners, and support what I can. I wish you well in your efforts.
Post reply on HN