Can anyone provide a reason an enterprise would choose Grok over a similar class of models?
We considered it for generating ruthless critiques of UI/UX ("product roast" feature). Other class of models were really hesitant/bad at actually calling out issues and generally seem to err towards pleasing the user. Here's a simple example I tried just now. Grok correctly removed mushrooms, but Chatgpt continues to try adding everything (I assume to be more compliant with the user): I only have pineapples, mushroom…
xAI's Grok 3 comes to Microsoft Azure
111–120 of 241 posts
Re: xAI's Grok 3 comes to Microsoft Azure
#112Honestly, Grok's technology is not impressive at all, and I wonder why anyone would use it: - Gemini is state-of-the-art for most tasks - ChatGPT has the best image generation - Claude is leading in coding solutions - Deepseek is getting old but it is open-source - Qwen has impressive lightweight models. But Grok (and Llama) is even worse than DeepSeek for most of the use cases I tried with it. The only thing it has…
Grok 3 mini is the best model in its price range for code, that doesn't train on your data. So it's part of Brokk's free plan. https://brokk.ai
Don't say that for sure unless you're inferencing it on your own machine.
Re: xAI's Grok 3 comes to Microsoft Azure
#113Honestly, Grok's technology is not impressive at all, and I wonder why anyone would use it: - Gemini is state-of-the-art for most tasks - ChatGPT has the best image generation - Claude is leading in coding solutions - Deepseek is getting old but it is open-source - Qwen has impressive lightweight models. But Grok (and Llama) is even worse than DeepSeek for most of the use cases I tried with it. The only thing it has…
Re: xAI's Grok 3 comes to Microsoft Azure
#114Earlier quoted context omitted.
Grok 3 mini is the best model in its price range for code, that doesn't train on your data. So it's part of Brokk's free plan. https://brokk.ai
> that doesn't train on your data. Don't say that for sure unless you're inferencing it on your own machine.
Re: xAI's Grok 3 comes to Microsoft Azure
#115Honestly, Grok's technology is not impressive at all, and I wonder why anyone would use it: - Gemini is state-of-the-art for most tasks - ChatGPT has the best image generation - Claude is leading in coding solutions - Deepseek is getting old but it is open-source - Qwen has impressive lightweight models. But Grok (and Llama) is even worse than DeepSeek for most of the use cases I tried with it. The only thing it has…
https://www.theguardian.com/technology/2025/may/14/elon-musk...
Re: xAI's Grok 3 comes to Microsoft Azure
#116Earlier quoted context omitted.
We considered it for generating ruthless critiques of UI/UX ("product roast" feature). Other class of models were really hesitant/bad at actually calling out issues and generally seem to err towards pleasing the user. Here's a simple example I tried just now. Grok correctly removed mushrooms, but Chatgpt continues to try adding everything (I assume to be more compliant with the user): I only have pineapples, mushroom…
What kind of test is that? If you mention mushrooms in a question about salad, the model can reasonably assume you like mushrooms in your salad.
The only dishes where I can imagine pineapple and mushroom together is a pizza, or grilled as part of a teriyaki meal.
Re: xAI's Grok 3 comes to Microsoft Azure
#117Honestly, Grok's technology is not impressive at all, and I wonder why anyone would use it: - Gemini is state-of-the-art for most tasks - ChatGPT has the best image generation - Claude is leading in coding solutions - Deepseek is getting old but it is open-source - Qwen has impressive lightweight models. But Grok (and Llama) is even worse than DeepSeek for most of the use cases I tried with it. The only thing it has…
At least two times they had unauthorized changes to their prompts to inject far right content that showed up on random content. imagine you're using it for a chat bot and it starts spouting off white nationalist content like "great replacement" theory. https://www.theguardian.com/technology/2025/may/14/elon-musk...
Re: xAI's Grok 3 comes to Microsoft Azure
#118It still seems to have the problems most other LLMs suffer with except Gemini: it loses context so quickly. I asked it about a paper I was looking at (SLOG [0]) and it basically lost the context of what "slog" referred to after 3 prompts. 1. I asked for an example transaction illustrating the key advantages of the SLOG approach. It responded with some general DB transaction stuff. 2. I then said "no use slog like we…
Re: xAI's Grok 3 comes to Microsoft Azure
#119Earlier quoted context omitted.
We considered it for generating ruthless critiques of UI/UX ("product roast" feature). Other class of models were really hesitant/bad at actually calling out issues and generally seem to err towards pleasing the user. Here's a simple example I tried just now. Grok correctly removed mushrooms, but Chatgpt continues to try adding everything (I assume to be more compliant with the user): I only have pineapples, mushroom…
What kind of test is that? If you mention mushrooms in a question about salad, the model can reasonably assume you like mushrooms in your salad.
Re: xAI's Grok 3 comes to Microsoft Azure
#120Honestly, Grok's technology is not impressive at all, and I wonder why anyone would use it: - Gemini is state-of-the-art for most tasks - ChatGPT has the best image generation - Claude is leading in coding solutions - Deepseek is getting old but it is open-source - Qwen has impressive lightweight models. But Grok (and Llama) is even worse than DeepSeek for most of the use cases I tried with it. The only thing it has…
Similarly I find grok is less likely to police itself to the point of retardation e.g. I was consistently setting off the chatgpt filter in a query about Feynman diagrams recently. Why?