The open weight model data is very interesting. I missed the release of Minimax M2. The benchmarks seem insanely impressive for its size. I would suspect benchmaxing but why would people be using it if it wasn’t useful?
State of AI: An Empirical 100T Token Study with OpenRouter
71–80 of 97 posts
Re: State of AI: An Empirical 100T Token Study with OpenRouter
#72I like to see stats like that, but I find it very concerning that OpenRouter don't mind inspecting its user/customer data without shame. Even if you pretend that the classifier respect anonymity, if I pay for the inference, I would expect that it would be a closed tube with my privacy respected. If at least it was for "safety" checks, I don't like that but I would almost understand, now it is for them to have "market…
Even better: they send all the data to GoogleTagClassifier, which means now Google had a copy of the sample as well
Re: State of AI: An Empirical 100T Token Study with OpenRouter
#73Earlier quoted context omitted.
Or maybe it’s just strange classification. I see a lot of prompts on the internet looking like “act as a senior xxx expert with over 15 years of industry experience and answer the following: [insert simple question]” I hope those are not classified as “roleplaying” the “roleplay” here is just a trick to get better answer from the model, often in a professional setting that has nothing to do with creative writing of N…
"act as a senior xxx expert with over 15 years of industry experience" ... I just don't get why LLMs are affected by this kind of nonsense -- is it due to training rewards?
Re: State of AI: An Empirical 100T Token Study with OpenRouter
#74According to the report, 52% of all open-source AI is used for *roleplaying*. They attribute it to fewer content filters and higher creativity. I'm pretty surprised by that, but I guess that also selects for people who would use openrouter
Re: State of AI: An Empirical 100T Token Study with OpenRouter
#75Most of the high volume enterprise use cases use their cloud providers (e.g., azure)
What we have here is mostly from smaller players. Good data but obviously a subset of the inference universe.
Re: State of AI: An Empirical 100T Token Study with OpenRouter
#76I like to see stats like that, but I find it very concerning that OpenRouter don't mind inspecting its user/customer data without shame. Even if you pretend that the classifier respect anonymity, if I pay for the inference, I would expect that it would be a closed tube with my privacy respected. If at least it was for "safety" checks, I don't like that but I would almost understand, now it is for them to have "market…
>I would expect that it would be a closed tube with my privacy respected Lol hahaha
"Everyone knows in {{BUSINESS_TYPE}} there is no real privacy".
Be it fintech, AI or social media. You give them a free pass with being flippant about companies respecting privacy.
Being flippant about anyone being careless about our privacy is doing us as a society and injustice. We should demand privacy, not laugh at the notion of privacy.
Re: State of AI: An Empirical 100T Token Study with OpenRouter
#77Super interesting data. I do question this finding: > the small model category as a whole is seeing its share of usage decline. It's important to remember that this data is from OpenRouter... a API service. Small models are exactly those that can be self-hosted. It could be the case that total small model usage has actually grown , but people are self-hosting rather than using an API. OpenRouter would not be in a pos…
Re: State of AI: An Empirical 100T Token Study with OpenRouter
#78According to the report, 52% of all open-source AI is used for *roleplaying*. They attribute it to fewer content filters and higher creativity. I'm pretty surprised by that, but I guess that also selects for people who would use openrouter
Or maybe it’s just strange classification. I see a lot of prompts on the internet looking like “act as a senior xxx expert with over 15 years of industry experience and answer the following: [insert simple question]” I hope those are not classified as “roleplaying” the “roleplay” here is just a trick to get better answer from the model, often in a professional setting that has nothing to do with creative writing of N…
> This indicates that users turn to open models primarily for creative interactive dialogues (such as storytelling, character roleplay, and gaming scenarios) and for coding-related tasks. The dominance of roleplay (hovering at more than 50% of all OSS tokens) underscores a use case where open models have an edge: they can be utilized for creativity and are often less constrained by content filters, making them attractive for fantasy or entertainment applications. Roleplay tasks require flexible responses, context retention, and emotional nuance - attributes that open models can deliver effectively without being heavily restricted by commercial safety or moderation layers. This makes them particularly appealing for communities experimenting with character-driven experiences, fan fiction, interactive games, and simulation environments.
I could imagine something like D&D or other types of narrative adventures on demand with a machine that never tires of exploring subplots or rewriting sections to be a bit different is a pretty cool thing to have. Either that, or writing fiction, albeit hopefully not entire slop books that are sold, but something to draw inspiration from and do a back and forth.
In regards to NSFW stuff, a while back people were clowning on OpenAI for suggesting that they'd provide adult writing content to adults, but it might as well be a bunch of money that's otherwise left on the table. Note: I'm all for personal freedom, though one also has to wonder about the longer term impact of those "AI girlfriend/boyfriend" trends, you sometimes see people making videos about those subreddits. Oh well, not my place to judge.
Edit: oh hey, there is more data there after all
> Among the highest-volume categories, roleplay stands out for its consistency and specialization. Nearly 60% of roleplay tokens fall under Games/Roleplaying Games, suggesting that users treat LLMs less as casual chatbots and more as structured roleplaying or character engines. This is further reinforced by the presence of Writers Resources (15.6%) and Adult content (15.4%), pointing to a blend of interactive fiction, scenario generation, and personal fantasy. Contrary to assumptions that roleplay is mostly informal dialogue, the data show a well-defined and replicable genre-based use case.
Re: State of AI: An Empirical 100T Token Study with OpenRouter
#79Earlier quoted context omitted.
Almost certainly VPN traffic. Most major LLMs block both China and Hong Kong (surprisingly, not the other way around), so Singapore ends up being the fastest nearby endpoint that isn't restricted.
Why do major LLMs block china? Isn't that a potentially huge market for them?
Re: State of AI: An Empirical 100T Token Study with OpenRouter
#80Earlier quoted context omitted.
Or maybe it’s just strange classification. I see a lot of prompts on the internet looking like “act as a senior xxx expert with over 15 years of industry experience and answer the following: [insert simple question]” I hope those are not classified as “roleplaying” the “roleplay” here is just a trick to get better answer from the model, often in a professional setting that has nothing to do with creative writing of N…
OpenRouter classifies content by the app that's used to interact with the llm. https://openrouter.ai/docs/app-attribution