Live data from Hacker News

Grok 4.6

x.ai

351–360 of 696 posts

Re: Grok 4.6

#351
post #129

In terms of using experience, I found Grok 4.5 to be way more pleasant to use than GPT 5.6 Sol and Claude 4.8/5. It just gets to the point, and is super fast and concise, no yapping. That's how AI agents should be imo. None of the weird "Claude ipsum" jargon like "load-bearing" and "stale folklore" or GPT 5.6-isms like "focused regression" and "provenance".

I’m not doing any coding with AI, so I’m the odd one out. Mostly use it for research: information retrieval and grokking technical concepts for exam prep. Does that fall within “knowledge work”?

Anywho - I switched to Opus last week and felt torn. It’s displayed somewhat higher competency in some responses, and the artifacts (diagrams) are splendid, but I despise its writing style. Grok is indeed fact/truth oriented, direct, and less personable (which I vastly prefer). Maybe I’ll switch back to Grok.

Re: Grok 4.6

#352
post #291

Earlier quoted context omitted.

[flagged]

So, what do you do for a living? You've made a personal attack and seem to be under the impression you're morally superior. So, I'm curious as to what highly virtuous role you take on in your daily life. That said, I see your comment history is a lot of one sentence personal attacks against people. Not a lot of thoughtful debate. This makes hypocrisy out of your supposed concern for social good.

I work for myself and I don't enable the production of CSAM so yeah I'm quite content being morally superior on this issue

Re: Grok 4.6

#353
post #45
post #23

Anyone else find it weird how within 2 months of Fable releasing all the major labs suddenly had Fable-level models? Trying to think of explanations: 1) AI researchers talk and change companies often, so techniques circulate. This feels implausible because training and shipping a new model ought to take longer than 2 months? 2) Distillation - also implausible for the reason above. 3) Benchmark hacking. AI companies h…

Yeah, I’m not convinced that there are any models as smart as Fable. Opus 5 definitely isn’t for all it has great benchmark scores. Fable displays judgement in a way I haven’t seen from any other model.

My experience with Fable is that it eats all my tokens and returns something I didn't ask for.

I realise this might be a skill issue.

I prefer models that are less "smart" but faster. Do the thing I asked you to do, immediately, and if you can't tell me and we'll work it through. Iterate faster not smarter.

Re: Grok 4.6

#354

As polarizing as grok is, it was basically inevitable for it to start being a real competitor given how much investment SpaceX made into its own inference capabilities. Seems if you are okay with it, there's no reason to use anything but the highest effort levels of some other frontier models for the price. I think Grok provides healthy competition to the other labs, though I do think they bank on groks reputation ma…

I can't bring myself to even try it. The guy did a salute on stage then spent billions of dollars on a mission to root out brown people who "didn't deserve" the position they were in. I feel gross just accidentally clicking links to x.

Re: Grok 4.6

#355
post #23

Anyone else find it weird how within 2 months of Fable releasing all the major labs suddenly had Fable-level models? Trying to think of explanations: 1) AI researchers talk and change companies often, so techniques circulate. This feels implausible because training and shipping a new model ought to take longer than 2 months? 2) Distillation - also implausible for the reason above. 3) Benchmark hacking. AI companies h…

benchamaxxing is easy, fable still seems organically more intelligent

Re: Grok 4.6

#356
post #53

As polarizing as grok is, it was basically inevitable for it to start being a real competitor given how much investment SpaceX made into its own inference capabilities. Seems if you are okay with it, there's no reason to use anything but the highest effort levels of some other frontier models for the price. I think Grok provides healthy competition to the other labs, though I do think they bank on groks reputation ma…

[flagged]

[flagged]

Re: Grok 4.6

#357

Earlier quoted context omitted.

It's a hack but doing things the 'proper' way is at least 1000x harder so whatever.

Is it? OpenAI released a gpt oss safeguard. You give it a policy it gives you a Rating Messages comes in rate it and reject with hitting the model. Then you don’t need to fill the prompt with “please don’t do this” https://huggingface.co/openai/gpt-oss-safeguard-120b

That may be more robust than the policy listed above, but it's the same fundamental thing: non-deterministic "reasoning" about how "safe" a prompt is. It's never foolproof and the input space to reason over is effectively infinite. You can only expect so much from prompts and models.

Re: Grok 4.6

#358
post #23

Anyone else find it weird how within 2 months of Fable releasing all the major labs suddenly had Fable-level models? Trying to think of explanations: 1) AI researchers talk and change companies often, so techniques circulate. This feels implausible because training and shipping a new model ought to take longer than 2 months? 2) Distillation - also implausible for the reason above. 3) Benchmark hacking. AI companies h…

I think it also shows that breakthroughs are not driven by innovative and research but mostly by scaling.

If this is the case, makes sense that frontier labs with similar access to compute driven by funding on same order of scale can produce improvement largely on similar pace

Re: Grok 4.6

#359

Earlier quoted context omitted.

> He's richer than that, think bigger. that's the hilarious paradox at the center of his antics. Musk is infamously petty and insecure. We're talking about the guy who tweaked Grok's system prompt to flatter him and paid someone to boost his fucking Diablo character for clout. I wouldn't put "looking through chat histories" past him for one second.

Reddits owner is also petty and insecure and edited other peoples posts, Elon hasn't done that yet. Didn't seem to stop reddit from getting popular, people don't really care that much.

...but people don't just hate Elon because he's petty? they hate him for the prejudiced BS and his actually harmful meddling in politics

Re: Grok 4.6

#360

I will say this: Grok Build has a very nice TUI! It even has... mouse rollovers/tooltips?? I was like whoa . I used Grok 4.5 for a security review the other day and it did a FANTASTIC job. I mean it thoroughly ROUTED my app's security, identifying attack surfaces I'd never even considered, and I LOVED it! (Guess why I had to use Grok to do the security review in the first place?!?! ) I'd suggest trying it out with so…

> mouse rollovers/tooltips?? I was like whoa

We're reinventing the wheel we tried to avoid in the first place.

Post reply on HN