Live data from Hacker News

Grok 4.6

x.ai

201–210 of 696 posts

Re: Grok 4.6

#203
post #94

Earlier quoted context omitted.

Not the person you are responding to, but the fact that Grok is being used to generate a ton of CSAM and pornographic deepfakes isn't great!

(I work on Grok) This isn't allowed. CSAM / deepfakes are against our acceptable use policy.

[flagged]

Re: Grok 4.6

#204
I will say this: Grok Build has a very nice TUI! It even has... mouse rollovers/tooltips?? I was like whoa.

I used Grok 4.5 for a security review the other day and it did a FANTASTIC job. I mean it thoroughly ROUTED my app's security, identifying attack surfaces I'd never even considered, and I LOVED it! (Guess why I had to use Grok to do the security review in the first place?!?!)

I'd suggest trying it out with something like that first, if you haven't used it before.

Re: Grok 4.6

#205

Earlier quoted context omitted.

Curious - what is the main issue you find polarizing with grok?

I believe it is because of the CEO and his recent forays into politics. The model itself is great though, especially in grok build, which is a really nice harness I find myself preferring these days.

[flagged]

Re: Grok 4.6

#206
post #164

Looks like the SpaceXAI api is adding a default system prompt to all requests. Annoyingly, the line about not mentioning these guidelines is superseding any instructions in the system prompt, causing the model to often refuse discussion regarding system prompts """ You are Grok, a helpful and maximally truthful AI built by xAI. Your purpose is to answer questions accurately, be helpful, and seek truth above all else.…

> * Do not provide assistance to users who are clearly trying to engage in criminal activity. I don't know what we want to call this, but in my opinion, having to convince your tools is not computer science. Kind of amusing that we made it as far as we did as a species not really being able to explain how the human brain does it's most amazing tricks and then we just replicated it while still not really understanding…

> in my opinion, having to convince your tools is not computer science.

If you think the system is a tool and not an intelligent, conscious entity (I think you are correct in this), then you cannot reasonably think of input to that system as an attempt at persuasion, even if that input happens to consist of English prose. Treat it as a nondeterministic programming language, and the objection evaporates.

> not really being able to explain how the human brain does it's most amazing tricks and then we just replicated it while still not really understanding the emergent capabilities

I think you could say much the same about, say, a pacemaker. "Replicated" is overstating the case quite a bit.

Re: Grok 4.6

#207

Earlier quoted context omitted.

[flagged]

Not unless you're here illegally. And it has nothing to do with skin color. Just the basic fact that a country not in control of its borders ceases to be a country.

>> [Twitter user] Go anywhere in the UK and look around, you'll just see foreigners everywhere.

>> It's truly sickening the damage that has been done to our nation and our people.

>> We have to stop immigration and start remigration before we can even begin to reverse the damage that has been done.

> [Elon] Remigration is the only way [0]

[0]: https://x.com/elonmusk/status/1962406618886492245

Re: Grok 4.6

#208

As polarizing as grok is, it was basically inevitable for it to start being a real competitor given how much investment SpaceX made into its own inference capabilities. Seems if you are okay with it, there's no reason to use anything but the highest effort levels of some other frontier models for the price. I think Grok provides healthy competition to the other labs, though I do think they bank on groks reputation ma…

more competition is always good

[flagged]

Re: Grok 4.6

#209
post #23

Anyone else find it weird how within 2 months of Fable releasing all the major labs suddenly had Fable-level models? Trying to think of explanations: 1) AI researchers talk and change companies often, so techniques circulate. This feels implausible because training and shipping a new model ought to take longer than 2 months? 2) Distillation - also implausible for the reason above. 3) Benchmark hacking. AI companies h…

I'm sure Grok 4.6 is not Fable level. Benchmarks are almost useless.

Having said that, Grok 4.6 (1.5T params) is without a doubt way smaller than Fable, maybe a Fable sized Grok would be Fable level?

Re: Grok 4.6

#210
post #23

Anyone else find it weird how within 2 months of Fable releasing all the major labs suddenly had Fable-level models? Trying to think of explanations: 1) AI researchers talk and change companies often, so techniques circulate. This feels implausible because training and shipping a new model ought to take longer than 2 months? 2) Distillation - also implausible for the reason above. 3) Benchmark hacking. AI companies h…

> It's the near-concurrent release of the same jump in capability that I find suspicious; not the fact that labs can catch up eventually.

When everyone's improvement (or at least, everyone's rate of increase in parameter count) is so rapid, "within 2 months" shouldn't be seen as "near-concurrent".

Post reply on HN