VibeThinker: 3B param model that beats Opus 4.5 on reasoning with novel SFT+GRPO
1–10 of 226 posts
Re: VibeThinker: 3B param model that beats Opus 4.5 on reasoning with novel SFT+GRPO
#2Re: VibeThinker: 3B param model that beats Opus 4.5 on reasoning with novel SFT+GRPO
#3Re: VibeThinker: 3B param model that beats Opus 4.5 on reasoning with novel SFT+GRPO
#4Re: VibeThinker: 3B param model that beats Opus 4.5 on reasoning with novel SFT+GRPO
#5I tried generating the classic pelican svg, but it failed horribly just showing me a rectangle and a black circle...
Re: VibeThinker: 3B param model that beats Opus 4.5 on reasoning with novel SFT+GRPO
#6I tried generating the classic pelican svg, but it failed horribly just showing me a rectangle and a black circle...
Re: VibeThinker: 3B param model that beats Opus 4.5 on reasoning with novel SFT+GRPO
#7Re: VibeThinker: 3B param model that beats Opus 4.5 on reasoning with novel SFT+GRPO
#8Re: VibeThinker: 3B param model that beats Opus 4.5 on reasoning with novel SFT+GRPO
#9I tried generating the classic pelican svg, but it failed horribly just showing me a rectangle and a black circle...
> these findings motivate the Parametric Compression-Coverage Hypothesis, which views verifiable reasoning as compressible into compact reasoning cores, while open-domain knowledge and general-purpose competence require broad parameter coverage over facts, concepts, and long-tail scenarios.
Re: VibeThinker: 3B param model that beats Opus 4.5 on reasoning with novel SFT+GRPO
#10Earlier quoted context omitted.
Its for reasoning not generating art?
Can you explain this a bit more
It would look really dumb if someone asked it that, but that's fine. You're trying to make a model that is optimized for efficiency for a specific task. As much as possible, you should prune uncorrelated things.