Live data from Hacker News

Grok 4.6

x.ai

91–100 of 696 posts

Re: Grok 4.6

#91

Earlier quoted context omitted.

Not the person you are responding to, but the fact that Grok is being used to generate a ton of CSAM and pornographic deepfakes isn't great!

Is that still a thing? I assumed they would have done something about it by now.

[flagged]

Re: Grok 4.6

#92
post #23

Anyone else find it weird how within 2 months of Fable releasing all the major labs suddenly had Fable-level models? Trying to think of explanations: 1) AI researchers talk and change companies often, so techniques circulate. This feels implausible because training and shipping a new model ought to take longer than 2 months? 2) Distillation - also implausible for the reason above. 3) Benchmark hacking. AI companies h…

I'm solidly in the "they are benchmaxxing" camp. This became very apparent with GPT 5.6 Sol. It, too, was widely hailed to have near-Fable level intelligence. But I used it non-stop for a week and realized that they had mostly just dialed up the relentlessness meter to eleven, most likely via heavy RLHF. Last week I gave it a small-sized auth ticket to work on, then stepped away. I came back later that afternoon and…

Yeah I found the timing on Sol especially curious since it came right on the heels of Fable. I've had mixed results with it - sometimes it seems great, other times it makes mistakes so stupid I cannot understand how it ever gets anything right.

Explaining it as a difference of effort would explain both.

Re: Grok 4.6

#93

>Grok 4.6 produces stronger first passes on visual and interactive projects than we typically saw with Grok 4.5. Given a concrete product idea, it is able to establish structure and visual language for an application in one pass. As a designer, I'm always hesitant to believe these statements until there's independent comparisons between the old & new model, as well as comparisons to human made flows. Design can be so…

(I work on Grok) We've been working on teaching the model how to reason about great visual design principles. Obviously this is hard and somewhat subjective, but through a combination of writing down these principles (e.g. how to think about systems, not just "use this italic serif font on marketing pages"), and then creating a lot of data to pairwise compare designs/outputs, we've made a notable improvement over G4.5 and see a path to improving much further in the next model.

Re: Grok 4.6

#94

Earlier quoted context omitted.

Curious - what is the main issue you find polarizing with grok?

Not the person you are responding to, but the fact that Grok is being used to generate a ton of CSAM and pornographic deepfakes isn't great!

(I work on Grok) This isn't allowed. CSAM / deepfakes are against our acceptable use policy.

Re: Grok 4.6

#95
post #23

Anyone else find it weird how within 2 months of Fable releasing all the major labs suddenly had Fable-level models? Trying to think of explanations: 1) AI researchers talk and change companies often, so techniques circulate. This feels implausible because training and shipping a new model ought to take longer than 2 months? 2) Distillation - also implausible for the reason above. 3) Benchmark hacking. AI companies h…

> It's the near-concurrent release of the same jump in capability that I find suspicious; not the fact that labs can catch up eventually.

What are suspicious of? If the timing is similar maybe just everyone already are of similar capabilities and got there at a similar time?

> Anyone else find it weird how within 2 months of Fable releasing all the major labs suddenly had Fable-level models?

It means Anthropic had no real moat and no real lead. Is that weird to you?

Re: Grok 4.6

#96
post #84

It's crazy that I'd literally trust a Chinese AI company with my data over anything Musk is involved with. Like, even if you don't care about (or even like) his politics and can look past how unlikable he comes off as, the damage he's done to his own reputation in this domain just makes using his products like this a no-go. He's literally so rich that he can get caught personally looking through chat sessions and it…

> It's crazy that I'd literally trust a Chinese AI company with my data It's crazy how much Chinese = bad the media or US companies have washed into you. Why lump it together? Like any place and any company there are good and bad 1s. It's not the Wild West over there...

I mean, the Chinese government doesn't really believe in checks and balances, or corporations as autonomous to the state. That's not a conspiracy, that's just how the CCP sees it (ask Jack Ma). You could argue the US has the Cloud Act, and obviously their respect for rules based law and order as a concept has heavily deteriorated, for but it's a very different kettle of fish to a regime who just doesn't even believe in the concept.

Re: Grok 4.6

#97
post #84

Earlier quoted context omitted.

> It's crazy that I'd literally trust a Chinese AI company with my data It's crazy how much Chinese = bad the media or US companies have washed into you. Why lump it together? Like any place and any company there are good and bad 1s. It's not the Wild West over there...

I mean, the Chinese government doesn't really believe in checks and balances, or corporations as autonomous to the state. That's not a conspiracy, that's just how the CCP sees it (ask Jack Ma). You could argue the US has the Cloud Act, and obviously their respect for rules based law and order as a concept has heavily deteriorated, for but it's a very different kettle of fish to a regime who just doesn't even believe…

Meanwhile Trump is building a surveillance state with all his tech executives friends who all massively benefit from government sponsored schemes, it's TOTALLY different!

Re: Grok 4.6

#98
post #23

Anyone else find it weird how within 2 months of Fable releasing all the major labs suddenly had Fable-level models? Trying to think of explanations: 1) AI researchers talk and change companies often, so techniques circulate. This feels implausible because training and shipping a new model ought to take longer than 2 months? 2) Distillation - also implausible for the reason above. 3) Benchmark hacking. AI companies h…

No, it has happened to almost every other "sota" model before. There used to be a meme with a circular arrow going through Anthropic, OpenAI, Google as a hype circle. Now we can drop Google and add a couple of Chinese companies.

It's not an explanation of why it happens, I am just pointing Fable is not an exception, it has happened with almost every other model release by all these companies over the last 2-3 years.

Re: Grok 4.6

#99
post #94

Earlier quoted context omitted.

Not the person you are responding to, but the fact that Grok is being used to generate a ton of CSAM and pornographic deepfakes isn't great!

(I work on Grok) This isn't allowed. CSAM / deepfakes are against our acceptable use policy.

Enforce it then

Re: Grok 4.6

#100

It's crazy that I'd literally trust a Chinese AI company with my data over anything Musk is involved with. Like, even if you don't care about (or even like) his politics and can look past how unlikable he comes off as, the damage he's done to his own reputation in this domain just makes using his products like this a no-go. He's literally so rich that he can get caught personally looking through chat sessions and it…

[flagged]
Post reply on HN