Do we have a good, objective benchmark set of prompts in existence somewhere? If not, I think having one would really help with tracking changes like that. I'm always skeptical of subjective feelings of tough-to-quantify things getting worse or better, especially where there is as much hype as for the various AI models. One explanation for the feelings is the model really getting significantly worse over time. Anothe…
Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?
271–280 of 817 posts
Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?
#272Earlier quoted context omitted.
It’s go-to tactic now if I ask it to go over any piece of code is to give a generic overview. Earlier, it would section out the code into chunks and go through each one individually.
Yeah, the bing integration did not go well. Went from amazing to annoying.
Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?
#273There's no doubt that it's gotten a lot worse on coding, I've been using this benchmark on each new version of GPT-4 "Write a tiptap extension that toggles classes" and so far it's gotten it right every time, but not any more, now it hallucinates a simplified solution that don't even use the tiptap api any more. It's also 200% more verbose in explaining it's reasoning, even if that reasoning makes no sense whatsoever…
Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?
#274My guess is that -probably no. It's more likely you had a stream of good luck in your earlier interactions and now you're observing regression to the mean. That can easily happen and it's why, for example, medical studies, are not taken as definitive proof of an effect. To further clarify, regression to the mean is the inevitable consequence of statistical error. Suppose (classic example) we want to test a hypertensi…
Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?
#275Earlier quoted context omitted.
So are you sure that isn't also nerfed?
It hasn't been changed since March 14th... So it's equally nerfed as it was then... Also, the playground lets you set the 'system message', which you can use to tell it to answer questions even if the results may be dangerous/rude/inappropriate.
Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?
#276Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?
#277OpenAI's models feel 100% nerfed to me at this point. I had it solving incredibly complex problems a few months ago (i.e. write a minimal PDF parser example), but today you will get scolded for asking such a complicated task of it. I think they programmed a classifier layer to detect certain coding tasks and shut it down with canned BS. I like to imagine certain billion/trillion-dollar mega corps had a back-room say…
Although I haven't had much of my time available for this recently. My recommendation would be to start with https://github.com/oobabooga/text-generation-webui
You will find almost everything you need to know there and on 4chan.org/g/catalog - search for LMG.
Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?
#278Yes. Before the update, when its avatar was still black, it solved pretty complex coding problems effortlessly and gave very nuanced, thoughtful answers to non-programming questions. Now it struggles with just changing two lines in a 10-line block of CSS and printing this modified 10-line block again. Some lines are missing, others are completely different for no reason. I'm sure scaling the model is hard, but they l…
"The original GPT-4 felt like magic to me" You never had access to that original. Watch this talk by one of the people that integrated GPT-4 in Bing telling how they noticed GPT-4 releases they got from OpenAI got iteratively and significantly nerfed even during the project. https://www.youtube.com/watch?v=qbIk7-JPB2c
Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?
#279The reason it's worse is basically because it's more 'safe' (not racist, etc). That of course sounds insane, and doesn't mean that safety shouldn't be strived for, etc - but there's an explanation as to how this occurs. It occurs because the system essentially does a latent classification of problems into 'acceptable' or 'not acceptable' to respond to. When this is done, a decent amount of information is lost regardi…
They're up against a pretty difficult barrier - if we had a perfect all-knowing oracle it might easily have opinions that are racist. Statistics alone suggest there will be racist truths. We're dealing with groups of people who are observably different from each other in correlated ways. GPT would need to reach a convincing balance of lying and honesty if it is supposed to navigate that challenge. It'd have to be dee…
How is stereotype different from pattern recognition?
These questions don't seem to go through the minds of people when developing "unbiased/impartial" technology.
There is no such thing as objective. So, why pretend to be objective and unbiased, when we all know its a lie?
Worst, if you pretend to be objective but aren't, then you are actually racist.
Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?
#280Earlier quoted context omitted.
Try out Bard, it's coding is much improved in the last 2 weeks. I've unfortunately switched over for the time being.
No thanks! I have better things to do than feeding that advertising behemoth. What I like about ChatGPT is that I don't see any ads at all!
Don't you worry, if there is any medium, place or mode of interaction people spend time on, advertising will eventually metastasize to it, and will keep growing until it completely devalues the activity and destroys most of the utility it provides.