News flash: people who scraped the Internet without permission to build their product complain when something vaguely similar is done to them. Water still wet, sky still blue. Film at 11. (slibhb: Don't get me wrong, I agree with you 99%. But the frontier labs have zero moral authority here.)
Qwen 3.8 follows GPT-5.5 Pro reasoning prefills
21–30 of 77 posts
Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills
#22This makes me very sad If Qwen and other Chinese labs are just copying reasoning traces, then those labs are more than a year behind the frontier.
Even if they weren't political adversaries of freedom, I would still feel 0% bad given all their training is already on data they got for free.
Information continues to want to be free. To the benefit of us all.
Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills
#23This makes me very sad If Qwen and other Chinese labs are just copying reasoning traces, then those labs are more than a year behind the frontier.
It makes me happy because it means that these misanthropic technofascists have no moat. They can spend trillions of dollars only for it to be largely copied in short order. Even if they weren't political adversaries of freedom, I would still feel 0% bad given all their training is already on data they got for free. Information continues to want to be free. To the benefit of us all.
Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills
#24Earlier quoted context omitted.
The thoughts trick was known before their paper / August. I "independently" "invented" it for the first Anthropic reasoning models because the API required you have thoughts for each assistant message. My app lets you switch AIs within a chat, and their API used to require thinking for all messages if thinking was enabled, so I needed to get a valid thinking stub to insert. Time has flew by for me the last 3 years, b…
Sure, but TFA is trying to use Qwen's reaction to the thoughts as proof that they did indeed extract thoughts to train on. My point is that any model trained after August 10 will know of those specific thoughts.
Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills
#25It's interesting that someone else noticed this. A week or two ago, GPT-5.6 Sol starting leaking reasoning into a tool call in Pi. I don't really know what happened, but it was ... interesting: Attach. Use hub debugger. Ensure source binary perhaps same. start. todo init. parallel no. two tool calls in same turn sequential is okay. immediately. exactly. Need not mention apologies yet final. [...] Let's do. [...] Do t…
What? I've never seen garbled CoT like the one you posted when using Qwen3.8-27B.
Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills
#26But for the question studied here it probably doesn't matter - overlaps in the publicly available output may be indicative of distillation (or not), regardless of what it is. I would just find it surprising that the Chinese labs would use it so trustingly. The publicly released reasoning trace is the first place where I would suspect some distillation poisoning to be injected.
Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills
#27This makes me very sad If Qwen and other Chinese labs are just copying reasoning traces, then those labs are more than a year behind the frontier.
Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills
#28This makes me very sad If Qwen and other Chinese labs are just copying reasoning traces, then those labs are more than a year behind the frontier.
Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills
#29It's interesting that someone else noticed this. A week or two ago, GPT-5.6 Sol starting leaking reasoning into a tool call in Pi. I don't really know what happened, but it was ... interesting: Attach. Use hub debugger. Ensure source binary perhaps same. start. todo init. parallel no. two tool calls in same turn sequential is okay. immediately. exactly. Need not mention apologies yet final. [...] Let's do. [...] Do t…
Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills
#30It's interesting that someone else noticed this. A week or two ago, GPT-5.6 Sol starting leaking reasoning into a tool call in Pi. I don't really know what happened, but it was ... interesting: Attach. Use hub debugger. Ensure source binary perhaps same. start. todo init. parallel no. two tool calls in same turn sequential is okay. immediately. exactly. Need not mention apologies yet final. [...] Let's do. [...] Do t…
> Qwen3.8 27B's CoT seems ... very similar. What? I've never seen garbled CoT like the one you posted when using Qwen3.8-27B.