Live data from Hacker News

xAI's Grok 3 comes to Microsoft Azure

techcrunch.com

81–90 of 241 posts

Re: xAI's Grok 3 comes to Microsoft Azure

#81
post #37

Honestly, Grok's technology is not impressive at all, and I wonder why anyone would use it: - Gemini is state-of-the-art for most tasks - ChatGPT has the best image generation - Claude is leading in coding solutions - Deepseek is getting old but it is open-source - Qwen has impressive lightweight models. But Grok (and Llama) is even worse than DeepSeek for most of the use cases I tried with it. The only thing it has…

The only interesting thing about Grok is using it hooked up to the X firehose to query about events in real time. Unfortunately it sucks at that.

Re: xAI's Grok 3 comes to Microsoft Azure

#82

Can anyone provide a reason an enterprise would choose Grok over a similar class of models?

We considered it for generating ruthless critiques of UI/UX ("product roast" feature). Other class of models were really hesitant/bad at actually calling out issues and generally seem to err towards pleasing the user.

Here's a simple example I tried just now. Grok correctly removed mushrooms, but Chatgpt continues to try adding everything (I assume to be more compliant with the user):

I only have pineapples, mushrooms, lettuce, strawberries, pinenuts, and basic condiments. What salad can I make that's yummy?

Grok: Pineapple-Strawberry Salad with Lettuce and Pine Nuts - https://x.com/i/grok/share/exvHu2ewjrWuRNjSJHkq7eLSY

ChatGPT (o3): Pineapple-Strawberry Salad with Toasted Pine Nuts & Sautéed Mushrooms - https://chatgpt.com/share/682b9987-9394-8011-9e55-15626db78b...

Re: xAI's Grok 3 comes to Microsoft Azure

#83

As someone developing agents using LLMs on various platform, im very reluctant to use anything associated with xAI. Grok's training data is increasingly pulled from an increasingly toxic source. Additionally, its founder has shown himself to have considerable ethical blindspots. Ive got enough second-order effects to be wary of. I cannot risk using technology with ethical concerns surrounding it as the foundation of…

>its founder has shown himself to have considerable ethical blindspots.

The guy is very vocal and clear about his ethical stances. Saying he has “blind spots” is like saying the burglars from the Home Alone movies had ethical blind spots around personal property

Re: xAI's Grok 3 comes to Microsoft Azure

#87
post #9

[flagged]

Ethics aside, we do not understand the technology enough to disentangle its outputs from the biases of its inputs. See the "Emergent misalignment" paper. The founder is clearly seeking to inject his ideology into this technology, so it is prudent to expect the technology to suffer in subtle and yet unidentified ways. This is Lysenkoism but for LLMs.

Re: xAI's Grok 3 comes to Microsoft Azure

#88
4-5 bn stake in xAI from Kingdom Holdings, who started Humain AI during the gulf tour. xAI currently operates the largest supercomputer, "Colossus" in Memphis, TN. Also, 5bn GB-200 NVIDIA server deal w/ Dell. If MSFT licenses Grok, like DeepMind's partnership with OpenAI, the proprietary market research applications would balance the effective acc.

Re: xAI's Grok 3 comes to Microsoft Azure

#89

Finally, I can use Microsoft's cloud to generate Zerohedge comments. > They also come with additional data integration, customization, and governance capabilities not necessarily offered by xAI through its API. Maybe we'll see a "Grok you can take to parties" come out of this.

Also, any other LLM is good for Reddit comments—-ironically.

Re: xAI's Grok 3 comes to Microsoft Azure

#90
post #82

Can anyone provide a reason an enterprise would choose Grok over a similar class of models?

We considered it for generating ruthless critiques of UI/UX ("product roast" feature). Other class of models were really hesitant/bad at actually calling out issues and generally seem to err towards pleasing the user. Here's a simple example I tried just now. Grok correctly removed mushrooms, but Chatgpt continues to try adding everything (I assume to be more compliant with the user): I only have pineapples, mushroom…

I haven't seen a model since the 3.5 Turbo days that can't be ruthless if asked to be. And Grok is about as helpful as any other model despite Elon's claims.

Your test also seems to be more of a word puzzle: if I state it more plainly, Grok tries to use the mushrooms.

https://grok.com/share/bGVnYWN5_2db81cd5-7092-4287-8530-4b9e...

And in fact, via the API with no system prompt it also uses mushrooms.

So like most models it just comes down to prompting.

Post reply on HN