Live data from Hacker News

Qwen 3.8 follows GPT-5.5 Pro reasoning prefills

gist.github.com

21–30 of 81 posts

Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills

#21

News flash: people who scraped the Internet without permission to build their product complain when something vaguely similar is done to them. Water still wet, sky still blue. Film at 11. (slibhb: Don't get me wrong, I agree with you 99%. But the frontier labs have zero moral authority here.)

[deleted]

Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills

#22

This makes me very sad If Qwen and other Chinese labs are just copying reasoning traces, then those labs are more than a year behind the frontier.

It makes me happy because it means that these misanthropic technofascists have no moat. They can spend trillions of dollars only for it to be largely copied in short order.

Even if they weren't political adversaries of freedom, I would still feel 0% bad given all their training is already on data they got for free.

Information continues to want to be free. To the benefit of us all.

Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills

#23

This makes me very sad If Qwen and other Chinese labs are just copying reasoning traces, then those labs are more than a year behind the frontier.

It makes me happy because it means that these misanthropic technofascists have no moat. They can spend trillions of dollars only for it to be largely copied in short order. Even if they weren't political adversaries of freedom, I would still feel 0% bad given all their training is already on data they got for free. Information continues to want to be free. To the benefit of us all.

Preach brother, they stole everything on the internet, and beyond, to train their models. They thought all that information was free, and everyone a few months beyond them is just following their example.

Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills

#24
post #19

Earlier quoted context omitted.

The thoughts trick was known before their paper / August. I "independently" "invented" it for the first Anthropic reasoning models because the API required you have thoughts for each assistant message. My app lets you switch AIs within a chat, and their API used to require thinking for all messages if thinking was enabled, so I needed to get a valid thinking stub to insert. Time has flew by for me the last 3 years, b…

Sure, but TFA is trying to use Qwen's reaction to the thoughts as proof that they did indeed extract thoughts to train on. My point is that any model trained after August 10 will know of those specific thoughts.

I'm sorry, it's going over my head still - my reading is "all models with any training after August 10 know how GPT 5.5 Pro thinks", but I'm not sure why - my initial guess was that's when GPT 5.5 was released, but that doesn't seem to be the case (it was released April 23rd).

Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills

#25
post #18

It's interesting that someone else noticed this. A week or two ago, GPT-5.6 Sol starting leaking reasoning into a tool call in Pi. I don't really know what happened, but it was ... interesting: Attach. Use hub debugger. Ensure source binary perhaps same. start. todo init. parallel no. two tool calls in same turn sequential is okay. immediately. exactly. Need not mention apologies yet final. [...] Let's do. [...] Do t…

> Qwen3.8 27B's CoT seems ... very similar.

What? I've never seen garbled CoT like the one you posted when using Qwen3.8-27B.

Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills

#26
I wasn't aware that we have access to raw reasoning tokens? I thought what you get is a kind of summary. Does the author have some kind of privileged access or was my assumption wrong?

But for the question studied here it probably doesn't matter - overlaps in the publicly available output may be indicative of distillation (or not), regardless of what it is. I would just find it surprising that the Chinese labs would use it so trustingly. The publicly released reasoning trace is the first place where I would suspect some distillation poisoning to be injected.

Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills

#27

This makes me very sad If Qwen and other Chinese labs are just copying reasoning traces, then those labs are more than a year behind the frontier.

It's been pretty obvious to me that the Chinese labs are operating mostly on a fast-follow strategy. The distillation attacks are well-documented, and there is good reason to believe they are able to copy architectural innovations as well. If US labs stagnate I would expect Chinese labs to stagnate as well. Their engineering is great, but in terms of frontier innovation (which requires heavy compute to search for new strategies that work at frontier scale) they are very far behind.

Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills

#28

This makes me very sad If Qwen and other Chinese labs are just copying reasoning traces, then those labs are more than a year behind the frontier.

The way I see it the Chinese labs are optimizing for other things, including effective compact models that don't need to run on top-of-the-line nVidia hardware.

Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills

#29
post #18

It's interesting that someone else noticed this. A week or two ago, GPT-5.6 Sol starting leaking reasoning into a tool call in Pi. I don't really know what happened, but it was ... interesting: Attach. Use hub debugger. Ensure source binary perhaps same. start. todo init. parallel no. two tool calls in same turn sequential is okay. immediately. exactly. Need not mention apologies yet final. [...] Let's do. [...] Do t…

I've read an embarrassing amount of Qwen 3.8 27b cot and it's nothing like this. I'm not refuting the OP, though, which is about continuation.

Re: Qwen 3.8 follows GPT-5.5 Pro reasoning prefills

#30
post #25
post #18

It's interesting that someone else noticed this. A week or two ago, GPT-5.6 Sol starting leaking reasoning into a tool call in Pi. I don't really know what happened, but it was ... interesting: Attach. Use hub debugger. Ensure source binary perhaps same. start. todo init. parallel no. two tool calls in same turn sequential is okay. immediately. exactly. Need not mention apologies yet final. [...] Let's do. [...] Do t…

> Qwen3.8 27B's CoT seems ... very similar. What? I've never seen garbled CoT like the one you posted when using Qwen3.8-27B.

I have seen plenty of Qwen 3.8 27B's caveman-like "Need doing this & that" thoughts. And on cerebras now I've seen them come real fast!
Post reply on HN