Live data from Hacker News

Grok 4.3

docs.x.ai

261–270 of 608 posts

Re: Grok 4.3

#261

Earlier quoted context omitted.

Elon publicaly claimed he had never corresponded with Epstein. that was a lie. When the documents were released they found several like thie one below. Saying things like "What day/night will be the wildest party on =our island?" [0] The "our" part is especially interesting as it implies he didnt just visit, but had an ownership stake. Other emails were found with Epstein making excuses to avoid having Musk visit, an…

My searches have not turned up a result showing that Musk "claimed he had never corresponded with Epstein". Can you source this? If not, can you explain why you did not check it before you posted the inaccurate claim?

At minimum Musk repeatedly claimed that Epstein was the one reaching out trying to get Musk to visit his island, when in reality Musk was the one initiating and asking which nights would be the wildest parties. And after making plans to visit with his then-wife, when Epstein warned him that the ratio of women-to-men might upset Musk’s wife, Musk told Epstein it wouldn’t be a problem.

https://www.theguardian.com/technology/2026/jan/30/elon-musk...

Musk has a long history of accusations (see the “I’ll buy you a horse” SpaceX lawsuit) as well as having fathered numerous children with women ~25 years younger than himself so not sure why you’d want to die on this particular hill.

Re: Grok 4.3

#262

How do the grok models fare in coding challenges to say gpt 5.5 and opus 4.6/4.7? I hate giving Elon any money. The man is a net negative to society but … if the models are objectively better then logically I must no?

All the downvotes are from Elon Stan’s. Think on your sins. ;-)

Re: Grok 4.3

#263
post #121

How do the grok models fare in coding challenges to say gpt 5.5 and opus 4.6/4.7? I hate giving Elon any money. The man is a net negative to society but … if the models are objectively better then logically I must no?

Logic can't tell you what your objectives should be, only how to achieve them.

Fair. Anyway I’ll look at benchmarks.

Re: Grok 4.3

#264

As an English-as-second-language speaker and writer, one thing Grok really shines at is capturing the tone and level of "formality" of a piece of text and the replicating it correctly. It seems to understand the little human subtleties of language in a way the other major providers don't. Chatgpt goes overly stiff and formal sounding, or ends up in a weird "aye guvnor" type informal language (Claude is sometimes bett…

I did a quick eval comparing Grok 4.3, Opus 4.7 and GPT 4.1 and they actually seem pretty similar: https://ofw640g9re.evvl.io/ They all did pretty well at a more "formal" tone, but GPT4.1 was the only one that didn't make me cringe with a "casual" tone. [edit] fwiw, grok was also the fastest+cheapest model, claude was slowest and priciest.

All three did well, and while I'm a Claude user, I found the Opus reply here added some unnecessary detail, like "Impact: Minimal; no downstream dependencies are currently at risk". Downstream dependencies weren't mentioned in the original message; for all we know downstream could be relying on a poorly performing API and is impacted by waiting another week for replacement.

Re: Grok 4.3

#265

Grok 4.3 was completed ahead of its CEO’s lesson on this common safety resource: Asked if he knew anything about OpenAI's "safety card," Musk smiled and replied: "Safety card? Why would it be a card?" https://www.axios.com/2026/04/30/musk-openai-safety-grok Low relevancy in spite of cluster size and musical chair gas generators for time being: Later in his testimony, Musk was asked about a claim he made last summer t…

Elon has publicly stated that he cares a great deal about safety. He has stated that the only safe models are those which align greatest with truth, that which is in reality. In this, xAI has lived up, as it has proved to hallucinate least (or close to least) in benchmarks.

If you read that, quote again, he is saying "how can you quantify safety in a card?"

Re: Grok 4.3

#266

So, we have: - claude for corps and gov - codex for devs - grok for what, roleplay, racism? Those are the two things I've ever heard grok associated with around me.

If you need to ask about what people on Twitter are talking about, Grok is really good for that obviously. I use it all the time for "what are the cool kids on twitter saying is the best tiling window manager these days" or whatever. Also, if you have a question that's borderline shady, Grok will often deliver. "Can you find a grey market Windows license site for me" etc.

btw copy pasted your idea in to supergrok, and learnt about Niri! Great use case, thanks!

Re: Grok 4.3

#267
post #36

So, we have: - claude for corps and gov - codex for devs - grok for what, roleplay, racism? Those are the two things I've ever heard grok associated with around me.

You should try all of them, then update your opinion about your information sources accordingly.

[flagged]

Re: Grok 4.3

#268
post #38

Earlier quoted context omitted.

I've also noticed that when I communicate with Grok in my native language, its tone is more natural than other models. I think this is due to the advantage of being trained on a large amount of Twitter data. However, as Twitter contains more and more AI-generated content now, I'm afraid continued training will make it less natural.

I'm sure Twitter knows which are the bot accounts and is surely excluding them from their model training. Twitter bots aren't a new phenomenon after all.

Highly doubtful seeing as my 14 year old twitter account got caught in a recent bot ban wave with no means of contacting a human for recovery.

Re: Grok 4.3

#269
post #206

While the tread is swapping between "OMG Claude good. OpenAI was done for" and "OMG Codex good. Anthropic was done for". I've never heard about Gemini and Grok. It works mostly similar performance, but people don't mention that much. Still, my impression is, Gemini hallucinate too much while Grok is always less capable than competitors so it's not worth using it.

Gemini 2.5 and 3 can code, but they are also dumb. They don't model the world well. It's hard to use them for programming tasks.

I haven't tried grok4.2 or grok4.3 yet for coding, but it wasn't up to the challenge as an agent yet. It looks like grok4.3 shifted its training and operates always as an agent first judging on some web usage. Musk knows grok is behind and states it publically. Now with grok4.3 release I do plan to try it again to see if it is suitable.

Re: Grok 4.3

#270
post #6

Ok speed (202.7 tok/s) and value (1.25 -> 2.50) look great, with pretty decent intelligence.

202.7 tok/s is only OK speed? Which providers are you using that are significantly better than that?

I said speed was great, Cerebas and Groq can provide better performance, likewise Fast versions of Cursor's Composer and Claude.

The reported speed like benchmarks is only a reported number on paper, we'll see how it holds up in real world usage, so far OpenRouter is only reporting 73tps

[1] https://openrouter.ai/x-ai/grok-4.3

Post reply on HN