Live data from Hacker News

Grok 4.5

x.ai

281–290 of 1001 posts

Re: Grok 4.5

#281

Earlier quoted context omitted.

You could be typing the same about Google or a number of the other labs right now. A diverse market full of choices keeps it from becoming the browser wars all over again.

> A diverse market full of choices keeps it from becoming the browser wars all over again. This is a great analogy but I worry you might be implying something I don't agree with but you didn't explicitly say what I'm worried about, so let me call it out: Microsoft played a dirty game with I.E, but they are in the dirty game business. It wasn't only I.E, it was their OS, Office suite and everything else they do busine…

"No one born in the LLM age even knows what I.E means or stands for, as it should be - a horribly designed, poorly working product"

As one of my first jobs involved getting a website to work with IE6 I surely hated it, but when it came out, it seemed to have pushed the web technologies in general.

The problem was not the browser technology, but microsoft abusing it's monopoly to don't give a shit about (open) web standards.

Re: Grok 4.5

#282
post #92

First impressions: - Very fast, easily beats GPT 5.5/Opus 4.8/GLM 5.2 because of higher t/s (around 90?) and very high token efficiency - Very good price, no contest vs GPT and Opus which are very overpriced if you pay API costs, and probably cheaper than GLM 5.2 when you take into account the token efficiency. - Will take quite a while to get a feel for how smart it is, but it's definitely good, I'd say in the same…

hmm interesting. maybe im doing something wrong. this model feels borderline unusable to me. it fumbles the most basic asks that require very little to no context consistently (inline these helper functions - re-rewrote half of the modules involved instead of making a 10 line change)

Re: Grok 4.5

#283

Can someone breakdown to me how this makes any sort of economical sense? Spending billions and billions to have the 3rd best model while even the number 1 and 2 players already seem to struggle making a profit. What am I missing here? Not trying to go full Ed Zitron but this doesn’t make sense to me.

„Markets can remain irrational longer than you can remain solvent.”

Re: Grok 4.5

#284
post #186
post #8

Of the 3 models I tried, Grok did the best at making an iOS app I wanted for personal use (a bike computer with specific qualities). (Claude just gave up and did an HTML/CSS implementation but I insisted on native SwiftUI+Metal.) Grok definitely fumbles sometimes, but I have been surprised what it CAN intuit versus me having to micromanage it. (I am not an iOS developer, so getting something specific that I needed in…

I swear I have read either this exact or a very similar comment before. Same gist about a bike computer iOS app, and one of the models giving up. As an aside, big thanks for Caddy! Really helped me get my greenfield project off the ground and it simply “just working” out of the box was one less source of errors I had to worry about when onboarding my team.

Wonderful, glad it was helpful for you!

Re: Grok 4.5

#285
post #75

With each release from the the other major labs, it becomes harder for Google to tell a compelling story about Gemini 3.5. Edit: Gemini 3.5 Pro . Expectations grow with each day it is not released.

Gemini is so far behind it hurts. It's useful for daily tasks and simple questions, but it codes like a model from late 2024. I can't imagine using it for any serious work.

In general I agree, but I found last week it was able to solve some obscure Android bugs for me that both 5.5 and Opus whiffed on.

Re: Grok 4.5

#286

Earlier quoted context omitted.

Composer 2.5 is 1T total/32B active (based on Kimi 2.5), while Elon publicly said Grok 4.5 is 1.5T parameters total. Hardly a different weight class. The API cost difference is ~2.5x, probably because xAI has much higher costs to recoup.

I could easily see Grok 4.5 being around 1:16 in terms of active parameters, so around 94B active parameters.

Why do you think that?

Re: Grok 4.5

#287

Earlier quoted context omitted.

Google is playing a different game. I don't really know what game they're playing, but they're not trying to beat Claude Code. They have coding capabilities and Antigravity, but I'd be surprised if it's much more than an afterthought. They're focusing on efficiency, models at the edge, human interaction, image and video, etc. in ways Anthropic, in particular, is not. Google wants its AI to be pervasive in everyone's…

At the same time that they’re seemingly exiting android?

And, yet. https://www.cnbc.com/2026/01/12/apple-google-ai-siri-gemini....

Re: Grok 4.5

#289

Earlier quoted context omitted.

Google is using AI at such scale internally they don't need external customers to recoup their investment.

> Google is using AI at such scale internally they don't need external customers to recoup their investment. That's assuming their flagship product remains relevant in an AI-powered world. Which brings to mind: most of the big shops product (chatgpt, claude, grok, etc...) ALL rely on search, and NONE of them actually have a running search stack. Which means, they must all be calling Google, no? How does Google make m…

> That's assuming their flagship product remains relevant in an AI-powered world.

The big advantage Google has, in my opinion, is Android. I think there is a decent chance that people stop downloading the ChatGPT, Claude, etc. apps if they perceive that the phone just does the same out of the box for free. And I reckon the majority of people will prefer free, ad-ridden AI chat vs. paying subscriptions, at least for personal use. And on the B2B side, they have Workspace deeply embedded in a huge number of companies. So I wouldn't count Google out.

Re: Grok 4.5

#290

[flagged]

Important part before parent comment gets dismissed:

> Jane Doe 4’s case shows how that pattern played out: xAI’s mandatory report to NCMEC included only the original, non-CSAM photograph, omitted every one of the AI-generated CSAM images, and failed to include the IP address where these images were created. Despite repeated requests from investigators for this location information that is critical for identifying and arresting perpetrators, xAI did not respond, stymieing the investigation for weeks.

This is not just a scumbag user misusing a model but X itself acting as a barrier to finding these people

Post reply on HN