Earlier quoted context omitted.
I think a more fair comparison would be that you hired a consulting firm to create a report and give you a summary of it, but you’re charged for the report itself separately from the summary, and you are not allowed to access the unsummarized report.
How is that a more fair comparison? The consulting firm in this case never promised you the interim reports, only the summaries of the reports. They also promised you the final output that the reports led to. You decided that report summaries + final output was worth paying for. You got exactly what you were promised.
Stealing Reasoning Traces from Proprietary LLM APIs
241–250 of 325 posts
Re: Stealing Reasoning Traces from Proprietary LLM APIs
#242Earlier quoted context omitted.
> But the real question is: Is it okay to steal By definition it cannot be stealing since you're paying for the tokens. It may be against their ToS, depending on what you end up doing with those tokens, but it cannot be stealing. If they charge by the token, all your tokens are belong to you :) I also find it very strange that everyone sort of accepts their ToS like no big deal. Imagine MS using the same terms for th…
> If they charge by the token, all your tokens are belong to you I’m not sure this argument is correct. You can sign whatever contract you like with the model provider, right? Including “you are entitled to the end product but not the intermediate scratch work”? Coming from a place of genuine curiosity: is there some precedent or statute that would invalidate that contract? I don’t see why the reasoning tokens belong…
Re: Stealing Reasoning Traces from Proprietary LLM APIs
#243Earlier quoted context omitted.
The only person calling it stealing is the author of this article, so this is a pointless discussion. The majority of this thread is just arguing with themselves.
They are also encrypting it so they must see some reason to do this. I suspect they think it is proprietary or otherwise a way that people can “steal” their implementations.
Re: Stealing Reasoning Traces from Proprietary LLM APIs
#244The hiding of this data only brings distrust to their frontier models. I think most people want to understand how something comes to a conclusion they don't want have that part left out on purpose...
it's this kind of behavior that forces people move to to open source models in the end, it's the lack of trust. the frontier model providers treat the end user/customer as a threat or adversary. Fable 5 is notorious for this. a lot of the serious questions you ask the model they won't even respond to you because of the woke guardrails. it wasn't only a couple weeks ago that huggingface had to use glm 5.2 to get the right answers about their security incident because Fable 5 didn't want to answer it.
Re: Stealing Reasoning Traces from Proprietary LLM APIs
#245Earlier quoted context omitted.
The reasoning blocks are not stolen/mined from the internet at large directly. They’re the result of a lot of research, time, money, and expertise into creating a reasoning model. To me the answer is quite clearly no, especially when the encrypted blocks demonstrate they want to protect it.
No. There are dozens of companies that resell tokens at a discount to collect and resell session data to various Chinese labs.
Re: Stealing Reasoning Traces from Proprietary LLM APIs
#246Re: Stealing Reasoning Traces from Proprietary LLM APIs
#247Earlier quoted context omitted.
> stop using morally charged terms made up by future monopolists Lets not gloss over this claim. Being: “Stealing is a morally charged term made up by future monopolists.” I strongly disagree. Stealing is not a made up term and property rights are foundational for any society. Your take is at least sensationalist if not malicious.
How much did anthropic pay to book publishers, and everybody else whose content they stole, for stealing their content ???
Re: Stealing Reasoning Traces from Proprietary LLM APIs
#248Earlier quoted context omitted.
How is that a more fair comparison? The consulting firm in this case never promised you the interim reports, only the summaries of the reports. They also promised you the final output that the reports led to. You decided that report summaries + final output was worth paying for. You got exactly what you were promised.
I personally read the thinking traces to know if the model is on the right direction
They explicitly do not promise reasoning traces. You (general you) agree to those terms and pay for that bargain anyways.
Re: Stealing Reasoning Traces from Proprietary LLM APIs
#249Apparently you can do the same by simply running it without reasoning, while giving it a thinking tool... >guys you do know you can just disable thinking, and instead give it a "deep_think" tool, and it will call it with internal CoT reasoning format right? >gl fixing that https://x.com/_can1357/status/2087228354399265125?s=20
Re: Stealing Reasoning Traces from Proprietary LLM APIs
#250"Stealing" something you already paid for (tokens), but that you can't have access to(!). And trained on the sum of human knowledge. Training on other model outputs ought to be business as usual, stop using morally charged terms made up by future monopolists: https://thomasdullien.github.io/posts/2026-06-15-rl-economic...
The only person calling it stealing is the author of this article, so this is a pointless discussion. The majority of this thread is just arguing with themselves.