Anyone else find it weird how within 2 months of Fable releasing all the major labs suddenly had Fable-level models? Trying to think of explanations: 1) AI researchers talk and change companies often, so techniques circulate. This feels implausible because training and shipping a new model ought to take longer than 2 months? 2) Distillation - also implausible for the reason above. 3) Benchmark hacking. AI companies h…
Possibility: They're all hitting the same plateau of what LLMs can do with their current architectures. I'm not stating this as a fact, but it's a hypothesis I'm keeping in my mix.
Grok 4.6
71–80 of 696 posts
Re: Grok 4.6
#72[flagged]
Re: Grok 4.6
#73Anyone else find it weird how within 2 months of Fable releasing all the major labs suddenly had Fable-level models? Trying to think of explanations: 1) AI researchers talk and change companies often, so techniques circulate. This feels implausible because training and shipping a new model ought to take longer than 2 months? 2) Distillation - also implausible for the reason above. 3) Benchmark hacking. AI companies h…
Other labs catching up in half a year seems about right.
Re: Grok 4.6
#74Fable-like intelligence, beats GPT-5.6-Sol on most benchmarks, cheaper than Kimi K3 on API and quite generous usage on Cursor subscription.
In my tests Grok 4.5 is definitely not Opus level. It is somewhere in between Sonnet and Opus, I'd say maybe a bit closer to Sonnet. We'll see with 4.6.
Re: Grok 4.6
#75As polarizing as grok is, it was basically inevitable for it to start being a real competitor given how much investment SpaceX made into its own inference capabilities. Seems if you are okay with it, there's no reason to use anything but the highest effort levels of some other frontier models for the price. I think Grok provides healthy competition to the other labs, though I do think they bank on groks reputation ma…
Curious - what is the main issue you find polarizing with grok?
The model itself is great though, especially in grok build, which is a really nice harness I find myself preferring these days.
Re: Grok 4.6
#76Earlier quoted context omitted.
Curious - what is the main issue you find polarizing with grok?
Not the person you are responding to, but the fact that Grok is being used to generate a ton of CSAM and pornographic deepfakes isn't great!
Re: Grok 4.6
#77[flagged]
Re: Grok 4.6
#78Anyone else find it weird how within 2 months of Fable releasing all the major labs suddenly had Fable-level models? Trying to think of explanations: 1) AI researchers talk and change companies often, so techniques circulate. This feels implausible because training and shipping a new model ought to take longer than 2 months? 2) Distillation - also implausible for the reason above. 3) Benchmark hacking. AI companies h…
> 1) AI researchers talk and change companies often, so techniques circulate. This feels implausible because training and shipping a new model ought to take longer than 2 months? The assumed timeline (2 months) is slightly wrong because Fable (Latin) is essentially the same as Mythos (Greek) albeit with protections against cyber and biological misuse. Mythos (Preview) was publicly announced in April 2026 [1] which me…
Re: Grok 4.6
#79Earlier quoted context omitted.
Curious - what is the main issue you find polarizing with grok?
I believe it is because of the CEO and his recent forays into politics. The model itself is great though, especially in grok build, which is a really nice harness I find myself preferring these days.
Re: Grok 4.6
#80Earlier quoted context omitted.
It's because Fable is just synthetic RL tasks + scale. The secret has been out for awhile now.
Does not explain timing