Live data from Hacker News

Grok 4.6

x.ai

291–300 of 696 posts

Re: Grok 4.6

#291
post #94

Earlier quoted context omitted.

Not the person you are responding to, but the fact that Grok is being used to generate a ton of CSAM and pornographic deepfakes isn't great!

(I work on Grok) This isn't allowed. CSAM / deepfakes are against our acceptable use policy.

[flagged]

Re: Grok 4.6

#292
I stopped bothering with Grok for anything when 4.5 dropped. It was so awful that I figured Elon had given up and was going to give alll his compute to Anthropic.

I’m extremely sceptical anyways - Grok 4.5 was probably the worst model I ever seriously tried to use going back 3 years.

Re: Grok 4.6

#293

Earlier quoted context omitted.

I use both Grok 4.5 and Opus 5. They’re both very good and Grok is faster and cheaper.

Opus 5 is terrible. I'd even say it's a step backwards from 4.8. I'm getting high error rates from it, and then it catches the error, and then it sometimes errors the error fix (!). Just today I had to switch another agent to Fable with the instruction, "Please clean up the mess that Opus 5 made, thanks" The other day, Sol called Opus 5's handoff (a skill I have that is basically a compaction, but just written to a f…

Thanks, that’s interesting to know. I don’t know much about LLMs so I use 5 because it’s a bigger number than 4.8.

Re: Grok 4.6

#294

Cursor blog: https://cursor.com/blog/grok-4-6

I'm a bit confused by the Cursor relationship here, the acquisition hasn't closed yet, what are they doing with Composer?

Re: Grok 4.6

#295

Earlier quoted context omitted.

I don't understand why they don't look for large substring matches for the system prompt before returning the response. Trivial calculation compared to a system prompt instruction asking the model not to do it

Because it's trivial to bypass through things like the model natively knowing how to speak in encodings like base64

Wait... Really!?

Re: Grok 4.6

#297

Earlier quoted context omitted.

I'd start here: https://en.wikipedia.org/wiki/Grok_(chatbot)#Controversies_a... And here: https://en.wikipedia.org/wiki/Grok_sexual_deepfake_scandal I think polarizing is a generous way of describing the problems. My organization has outright banned Grok, because we don't trust SpaceX to hold up to contractual agreements vis-a-vis data-privacy/training. That's the level of reputational damage we're talking about here…

Can someone help me understand the deep fake controversy? That's like making photoshop illegal.

[deleted]

Re: Grok 4.6

#299
post #43

Tangental, but has anyone else noticed grok's voice mode got stupid and terse ~2 weeks ago? I've absolutely loved grok's voice mode since it came out (incredibly useful for brainstorming on walks and helping conceptualise and get the verbiage for expressing ideas) but it seems so have lost about 40 IQ points recently, and if the question is multi-part, it often answers just one part with no elaboration or explanation…

I haven't experienced a regression, but voice modes have always been stupider than frontier models. In my experience Grok's voice mode suffers the least from this, and it's been getting better over time. It's especially good (compared to ChatGPT or Gemini) on things that involve current events or web research. Just yesterday in the car I got it to locate and read and explain a recent academic paper and multiple of my questions were answered with several minute long monologues that contained useful and accurate information.

Re: Grok 4.6

#300

Earlier quoted context omitted.

I'd start here: https://en.wikipedia.org/wiki/Grok_(chatbot)#Controversies_a... And here: https://en.wikipedia.org/wiki/Grok_sexual_deepfake_scandal I think polarizing is a generous way of describing the problems. My organization has outright banned Grok, because we don't trust SpaceX to hold up to contractual agreements vis-a-vis data-privacy/training. That's the level of reputational damage we're talking about here…

The US govt trusts SpaceXAI for defense and high security missions. The idea they are lying about contracted AI services is absurd. They're also a public company which beings even more oversight than openai / anthropic.

Running separate services for the government is very common in software services. Being public doesn't bring any technical oversight at all. I haven't actually heard of grok being used for the government security ive only ever heard Claude being used.
Post reply on HN