Live data from Hacker News

GPT-4.5

openai.com

571–580 of 1001 posts

Re: GPT-4.5

#571

Earlier quoted context omitted.

Huh. Disregarding the 4.5-specific bit here, a browser extension or possibly website that did this in general could be really useful. Maybe even something that just noticed whenever you visited a site that had had significant HN discussion in the past, then let you trigger a summary.

there are literally hundreds of extensions and sites that do this the problem is that they are competing each other into the ground hence they go unmaintained very quickly getrecall.ai has been the most mature so far

Thanks, it's amazing how much stuff is out there I don't know about.

Re: GPT-4.5

#572

Earlier quoted context omitted.

And LLM's already have tons of productive uses. The biggest ones are probably still waiting, though. But this is about one particular price/performance ratio. You need to build things before you can see how the market responds. You say it's "not good business" but that's entirely wrong. It's excellent business. It's the only way to go about it, in fact. Finding product-market fit is a process. Companies aren't omnisc…

> And LLM's already have tons of productive uses. I disagree strongly with that. Right now they are fun toys to play with, but not useful tools, because they are not reliable. If and when that gets fixed, maybe they will have productive uses. But for right now, not so much.

I use LLMs everyday to proofread and edit my emails. They’re incredible at it, as good as anyone I’ve ever met. Tasks that involve language and not facts tend to be done well by LLMs.

Re: GPT-4.5

#573

Earlier quoted context omitted.

> We look forward to learning more about its strengths, capabilities, and potential applications in real-world settings. If GPT‑4.5 delivers unique value for your use case, your feedback (opens in a new window) will play an important role in guiding our decision. "We don't really know what this is good for, but spent a lot of money and time making it and are under intense pressure to announce new things right now. If…

Maybe if they build a few more data centers, they'll be able to construct their machine god. Just a few more dedicated power plants, a lake or two, a few hundred billion more and they'll crack this thing wide open. And maybe Tesla is going to deliver truly full self driving tech any day now. And Star Citizen will prove to have been worth it along along, and Bitcoin will rain from the heavens. It's very difficult to r…

Correction: We're expected to pay for the ride, whether we choose to come along or not.

Re: GPT-4.5

#574

Earlier quoted context omitted.

> We don't really know what this is good for Oh come on. Think how long of a gap there was between the first microcomputer and VisiCalc. Or between the start of the internet and social networking. First of all, it's going to take us 10 years to figure out how to use LLM's to their full productive potential. And second of all, it's going to take us collectively a long time to also figure out how much accuracy is neces…

ChatGPT had its initial public release November 30th, 2022. That's 820 days to today. The Apple II was first sold June 10, 1977, and Visicalc was first sold October 17, 1979, which is 859 days. So we're right about the same distance in time- the exact equal duration will be April 7th of this year. Going back to the very first commercially available microcomputer, the Altair 8800 (which is not a great match, since tha…

> Visicalc was first sold October 17, 1979, which is 859 days.

And it still can't answer simple English-language questions.

Re: GPT-4.5

#575
post #518

Earlier quoted context omitted.

At least so far it's coding performance is bad, but from what I have seen it's writing abilities are totally insane. It doesn't read like AI output anymore.

Any examples you’d be willing to share?

They have examples in the announcement post. It does a better job of understanding intent in the question which helps it give an informal rather than essay style response where appropriate.

Re: GPT-4.5

#576

GPT pro already has already rummored to be 100k users. You think GPT 4.5 will add to that even with the insane costs for corporate users?

What rumors? I looked and can’t find something to substantiate that #

in my experience, o3-mini-high while still unpredictable as it modifies and ignores parts of my code when I specifically tell it not to do so (e.g. "don't touch anything else!") is the best AI coding tool out there, far better than Claude

so Pro is worth it for O3-mini-high

Re: GPT-4.5

#578
post #469
post #146

Earlier quoted context omitted.

Sam tweeted that they're running out of computer. I think it's reasonable to think they may serve somewhat quantized models when out of capacity. It would be a rational business decision that would minimally disrupt lower tier ChatGPT users. Anecdotally, I've noticed what appears to be drops in quality, some days. When the quality drops, it responds in odd ways when asked what model it is.

I mean, GPT 4.5 says "I'm ChatGPT, based on OpenAI's GPT-4 Turbo model." and o1 Pro Mode can't answer, just says "I’m ChatGPT, a large language model trained by OpenAI." Asking it what model it is shouldn't be considered a reliable indicator of anything.

> Asking it what model it is shouldn't be considered a reliable indicator of anything.

Sure, but a change in response may be, which is what I see (and no, I have no memories saved).

Re: GPT-4.5

#580
post #26

GPT 4.5 pricing is insane: Price Input: $75.00 / 1M tokens Cached input: $37.50 / 1M tokens Output: $150.00 / 1M tokens GPT 4o pricing for comparison: Price Input: $2.50 / 1M tokens Cached input: $1.25 / 1M tokens Output: $10.00 / 1M tokens It sounds like it's so expensive and the difference in usefulness is so lacking(?) they're not even gonna keep serving it in the API for long: > GPT‑4.5 is a very large and comput…

> We look forward to learning more about its strengths, capabilities, and potential applications in real-world settings. If GPT‑4.5 delivers unique value for your use case, your feedback (opens in a new window) will play an important role in guiding our decision. "We don't really know what this is good for, but spent a lot of money and time making it and are under intense pressure to announce new things right now. If…

Really don’t understand what’s the use case for this. The o series models are better and cheaper. Sonnet 3.7 smokes it on coding. Deepseek R1 is free and does a better job than any of OAI’s free models
Post reply on HN