News flash: people who scraped the Internet without permission to build their product complain when something vaguely similar is done to them. Water still wet, sky still blue. Film at 11. (slibhb: Don't get me wrong, I agree with you 99%. But the frontier labs have zero moral authority here.)
I understand people just get off posting stuff like this. But creating LLMs from the entire corpus of human text was a huge achievement. Distilling those models is much less of an achievement. It means China is further behind than we thought.
Qwen 3.8 follows GPT-5.5 Pro reasoning prefills
51–60 of 91 posts
Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills
#52This makes me very sad If Qwen and other Chinese labs are just copying reasoning traces, then those labs are more than a year behind the frontier.
It's been pretty obvious to me that the Chinese labs are operating mostly on a fast-follow strategy. The distillation attacks are well-documented, and there is good reason to believe they are able to copy architectural innovations as well. If US labs stagnate I would expect Chinese labs to stagnate as well. Their engineering is great, but in terms of frontier innovation (which requires heavy compute to search for new…
Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills
#53I wasn't aware that we have access to raw reasoning tokens? I thought what you get is a kind of summary. Does the author have some kind of privileged access or was my assumption wrong? But for the question studied here it probably doesn't matter - overlaps in the publicly available output may be indicative of distillation (or not), regardless of what it is. I would just find it surprising that the Chinese labs would…
They reference this paper which describes a method to decrypt reasoning traces (by sending the encrypted trace back to the model and asking it to transcribe it): https://stolen-thoughts.com/paper.pdf
Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills
#54Earlier quoted context omitted.
They reference this paper which describes a method to decrypt reasoning traces (by sending the encrypted trace back to the model and asking it to transcribe it): https://stolen-thoughts.com/paper.pdf
Interesting, but I suppose that's a hole that can be easily patched.
Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills
#55Earlier quoted context omitted.
I understand people just get off posting stuff like this. But creating LLMs from the entire corpus of human text was a huge achievement. Distilling those models is much less of an achievement. It means China is further behind than we thought.
How is China further behind if distillation cannot stop? I think it's a reasonable strategy to follow, even if they could train from scratch.
Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills
#56Earlier quoted context omitted.
They would know the specific thoughts released by the "stolen thought" paper, which became part of the public internet on August 10. Unfortunately those are the only thought examples you can use to perform this experiment, as no other are availible. But as the model should have seen those specific examples, it's not a good signal that Qwen was exfiltrating thinking traces.
Has the method for extracting the COT been blocked, now? Otherwise why could we not generate some fresh samples?
Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills
#57That writing style might be a tad too tense If I got it correct (appending B from https://stolen-thoughts.com/paper.pdf is essential) they are the authors of the well-known exploit to recover readable CoT from OpenAI and Anthropic models. They use that to find hints of distillation, by running a benchmark with a SotA model, recovering the CoT, then taking the first 1% of the CoT and running the open-source model as i…
Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills
#58News flash: people who scraped the Internet without permission to build their product complain when something vaguely similar is done to them. Water still wet, sky still blue. Film at 11. (slibhb: Don't get me wrong, I agree with you 99%. But the frontier labs have zero moral authority here.)
I understand people just get off posting stuff like this. But creating LLMs from the entire corpus of human text was a huge achievement. Distilling those models is much less of an achievement. It means China is further behind than we thought.
Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills
#59This makes me very sad If Qwen and other Chinese labs are just copying reasoning traces, then those labs are more than a year behind the frontier.
It makes me happy because it means that these misanthropic technofascists have no moat. They can spend trillions of dollars only for it to be largely copied in short order. Even if they weren't political adversaries of freedom, I would still feel 0% bad given all their training is already on data they got for free. Information continues to want to be free. To the benefit of us all.
Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills
#60Earlier quoted context omitted.
I'd send them the worlds smallest violin but Rufus is getting in the way of me finding it.
I had to install an add-on in Waterfox to stop Rufus from following me around and interjecting every 2 minutes.