Live data from Hacker News

GPT-5.2

openai.com

701–710 of 1001 posts

Re: GPT-5.2

#701
post #560

Earlier quoted context omitted.

Why no grok 4.1 reasoning?

Do people other than Elon fans use grok? Honest question. I've never tried it.

I dislike Musk, and use Grok. I find it most useful for analyzing text to help check if there's anything I've missed in my own reading. Having it built in to Twitter is convenient and it has a generous free tier.

Re: GPT-5.2

#703

I have been using chatGPT a ton over the last months and paying the subscription. Used it for coding, news, stock analysis, daily problems, and a whatever I could think of. I decided to give Gemini a go when version three came out to great reviews. Gemini handles every single one of my uses cases much better and consistently gives better answers. This is especially true for situations were searching the web for curre…

I see a post like this every time there are news about ChatGPT or OpenAI. I'm probably being paranoid but I keep thinking that it looks like bots or paid advertisement for Gemini

I think people like me just enjoying sharing when something is working for them and they have a good experience. It probably gets voted up because people enjoy reading when that happens

Re: GPT-5.2

#704

I have been using chatGPT a ton over the last months and paying the subscription. Used it for coding, news, stock analysis, daily problems, and a whatever I could think of. I decided to give Gemini a go when version three came out to great reviews. Gemini handles every single one of my uses cases much better and consistently gives better answers. This is especially true for situations were searching the web for curre…

Get Gemini answer and tell ChatGPT this is what my friend said. Then put ChatGPT answer to Claude and so on. It's a cheat code.

I did this today it was amazing. If I would have had time I would try other models as well. Great tip thanks

Re: GPT-5.2

#705
post #581

Earlier quoted context omitted.

I have tried the models and in domains I know well they are pathetic. They remove all nuance, make errors that non-experts do not notice and generally produce horrible code. It is even worse in non-programming domains, where they chop up 100 websites and serve you incorrect bland slop. If you are using them as a search helper, that sometimes works, though 2010 Google produced better results. Oracle dropped 11% today…

> they remove all nuance Said in a sweeping generalization with zero sense of irony :D

This is a good point. It is a sweeping generalization if you do not read the sentence that comes before that quote

Re: GPT-5.2

#706

So the rosy biased estimate is OpenAI is saving 1 hour of work per day, so 5 hours total per-work week and 20 hours total per-month. With a subsidized cost of $200/month for OpenAI it would be cheaper to hirer a part-time minimum wage worker than it would be to contract with OpenAI. And that is the rosiest estimate OpenAI has.

The closest I come to working with part-time, minimum-wage workers is working with student employees. Even then, they earn more and usually work more than five hours a week.

Most of the time, I end up putting in more work than I get out of it. Onboarding, reviewing, and mentoring all take significant time.

Even with the best students we had, paying around 400 euros a month, I would not say that I saved five hours a week.

And even when they reach the point of being truly productive, they are usually already finished with their studies. If we then hire them full-time, they cost significantly more.

Re: GPT-5.2

#707

Again I just tap the sign. All of your benchmarks mean nothing to me until you include Claude Sonnet on them. In my experience, GPT hasn’t been able to compete with Claude in years for the daily “economically valuable” tasks I work on.

Claude is pretty trash for anything besides coding

What are you basing that on? Between Sonnet and Opus I don't think I'm reaching for Gemini 3 at all.

Re: GPT-5.2

#708

I have been using chatGPT a ton over the last months and paying the subscription. Used it for coding, news, stock analysis, daily problems, and a whatever I could think of. I decided to give Gemini a go when version three came out to great reviews. Gemini handles every single one of my uses cases much better and consistently gives better answers. This is especially true for situations were searching the web for curre…

Ditto but for Claude -- blows GPT out of the water. Much better in coding and solving physics problems from the images (in foreign languages). GPT couldn't even read the image. The only annoying thing is that if you use Opus for coding, your usage will fill up pretty fast.

anyway, cancelled my chatgpt subscription.

Re: GPT-5.2

#709

Earlier quoted context omitted.

have been on 1M context window with claude since 4.0 - it gets pretty expensive when you run 1M context on a long running project (mostly using it in cline for coding). I think they've realized more context length = more $ when dealing with most agentic coding workflows on api.

You should be doing everything you can to keep context under 200k, ideally even 100k. All the models unwind so badly as context grows.

I don't have that experience with gemini. Up to 90% full, it's just fine.
Post reply on HN