Live data from Hacker News

State of AI: An Empirical 100T Token Study with OpenRouter

openrouter.ai

81–90 of 97 posts

Re: State of AI: An Empirical 100T Token Study with OpenRouter

#81
post #25

> The noticeable spike [~20 percentage points] in May in the figure above [tool invocations] was largely attributable to one sizable account whose activity briefly lifted overall volumes. The fact that one account can have such a noticeable effect on token usage is kind of insane. And also raises the question of how much token usage is coming from just one or five or ten sizeable accounts.

It is quite interesting to ponder these usage statistics, isn't it? According to their charts they're at a throughput of something like 7T tok/week total now. At 1$/Mtok, that's 7M$ per week. Less than half a billion per year. How much is that compared to the total inference market? And yet again, their throughput went like 20x in one year, who knows what's to come...

Yes, but that token growth chart looks linear to me. There's the usual summer slump and then growth catches up once the autumn begins, but if you plot a line from the winter growth period at the start of 2025 you end up roughly in the right place except for an unusual spike in the most recent month (maybe another big user).

I'd have liked to see a chart of all tokens broken down by category rather than just percentages, but what this data seems to be saying is that growth isn't exponential, and is being dominated by growth in programming. A lot of the spending in AI is being driven by the assumption that it'll be used for everything everywhere. Perhaps it's just OpenRouter's user base, but if this data is representative then it implies AI adoption isn't growing all that fast outside of the tech industry (especially as "science" is nearly all AI related discussion).

This feels intuitively likely. I haven't seen many obvious signs of AI adoption around me once I leave the office. Microsoft has been struggling to sell its Copilot offerings to ordinary MS Office users, who apparently aren't that keen. The big wins are going to be existing apps and data pipelines calling out to AI, and it'll just take time to figure out what those use cases are and integrate them. Integrating even present-day AI into the long tail of non-tech industries is probably going to take decades.

Also odd: no category for students cheating on homework? I notice that "editing services" is a big chunk of the "academia" category. Probably most of that traffic goes direct to chatgpt.com and bypasses OpenRouter entirely.

Re: State of AI: An Empirical 100T Token Study with OpenRouter

#82

I am a person who wants to maintain a distance from the AI-hype train, but seeing a chart like this [1], I can't help think that we are nowhere near the peak. The weekly token consumption keeps on rising, and it's already in trillions, and this ignores a lot of consumption happening directly through APIs. Nvidia could keep delivering record-breaking numbers, and we may well see multiple companies hit six, seven, or e…

Problem is that most of that growth is in models being underpriced. We don't know what the demand curve looks like when tokens are priced to cover the full operating costs of the companies making them.

Also, growth seems to be linear, not exponential.

Re: State of AI: An Empirical 100T Token Study with OpenRouter

#83

Earlier quoted context omitted.

>And then they reveal that they did actually have access to the text I'm not seeing that. All I'm seeing is them analyzing metadata.

>All I'm seeing is them analyzing metadata Read the section about how they achieve classifications for prompts (hint: They read the prompts)

I didn’t read the paper but I know OR has an option to opt-in to reading/training off the prompts for a discount. Some free models also log, but I’m not sure if that is just the provider, or OR too

Re: State of AI: An Empirical 100T Token Study with OpenRouter

#84
post #52

Earlier quoted context omitted.

They explicitly give you a discount if you opt in to allowing your data to be used for (anonymized) analytics. That’s pretty fair imho.

Cynical take: they could look at everyone and give a discount for optics. I'd feel a lot better if "OpenRouter" were open source.

litellm is basically open source version https://www.litellm.ai although the openrouter being a hosted service is kinda the point. Unless the whole industry decides to this over e2ee you cant get any guarantees about an intermediary aggregator

Re: State of AI: An Empirical 100T Token Study with OpenRouter

#85
post #8

According to the report, 52% of all open-source AI is used for *roleplaying*. They attribute it to fewer content filters and higher creativity. I'm pretty surprised by that, but I guess that also selects for people who would use openrouter

I'm not surprised at all. The HN crowd think LLMs are mostly used for engineering because they live in a multi layer bubble. Real people in the real world do all kind of shit with LLMs which aren't productivity or even work related.

Re: State of AI: An Empirical 100T Token Study with OpenRouter

#86

Earlier quoted context omitted.

>And then they reveal that they did actually have access to the text I'm not seeing that. All I'm seeing is them analyzing metadata.

>All I'm seeing is them analyzing metadata Read the section about how they achieve classifications for prompts (hint: They read the prompts)

From what I see the researchers aren't running a classifier on prompts they've acquired.

>The classifier is deployed within OpenRouter's infrastructure, ensuring that classifications remain anonymous and are not linked to individual customers.

OpenRouter has to have access to your prompts in order to route it somewhere else. The researchers don't get access to these prompts. They only get access to the metadata being generated from routing a prompt.

Re: State of AI: An Empirical 100T Token Study with OpenRouter

#87
post #52

Earlier quoted context omitted.

They explicitly give you a discount if you opt in to allowing your data to be used for (anonymized) analytics. That’s pretty fair imho.

Cynical take: they could look at everyone and give a discount for optics. I'd feel a lot better if "OpenRouter" were open source.

Nobody is forcing you to use it. It’s a service for convenience: you just pay one provider instead of having to create a bazillion accounts.

If you don’t like any middle men, just go to one of the providers directly.

Re: State of AI: An Empirical 100T Token Study with OpenRouter

#88

I like to see stats like that, but I find it very concerning that OpenRouter don't mind inspecting its user/customer data without shame. Even if you pretend that the classifier respect anonymity, if I pay for the inference, I would expect that it would be a closed tube with my privacy respected. If at least it was for "safety" checks, I don't like that but I would almost understand, now it is for them to have "market…

> Imagine

why imagine? The world already functions exactly like that. Talk on Tg like every chat is summarized every 24hrs and monthly (with cheap LLM and then with strong ones if signals found), and it reports to all kinds of interested intel agencies.

Same for openrouter. everything that leaves your device plaintext = public. Period. No hopes.

Re: State of AI: An Empirical 100T Token Study with OpenRouter

#89
The 4x growth in prompt length is a fundamental shift. We've quickly moved from "Q&A" mode to "upload full context and analyze" mode.

This completely changes infrastructure requirements: KV-caching becomes a necessity, and prefill time becomes a critical metric, often more important than generation speed. That's exactly why models with cheap long context (Gemini, DeepSeek) are winning the race against "smarter" but expensive models. Inference economics are now dictated by context length

Re: State of AI: An Empirical 100T Token Study with OpenRouter

#90

Earlier quoted context omitted.

>I would expect that it would be a closed tube with my privacy respected Lol hahaha

Comments such as these are what allows developers to justify ignoring our privacy. "Everyone knows in {{BUSINESS_TYPE}} there is no real privacy". Be it fintech, AI or social media. You give them a free pass with being flippant about companies respecting privacy. Being flippant about anyone being careless about our privacy is doing us as a society and injustice. We should demand privacy, not laugh at the notion of pr…

The parent comment is exaclty right. "LOL" is best possible response.

>We should demand privacy, not laugh at the notion of privacy.

Recently got m3 ultra 512gb studio. LM Studio runs frontier models routinely. Going local is the ONLY way. That's all you can do. "Demanding privacy" is security theater. Act accordingly.

Post reply on HN