Is it consistently worse or just sometimes/often worse than before? Any extreme power users or GPT-whisperers here? If it’s only noticeably worse X% of the time my bet would be experimentation. One of my least favorite patterns that tech companies do is use “Experimentation” overzealously or prematurely. Mainly, my problem is they’re not transparent about it, and it creates an inconsistent product experience that jus…
No place I worked at ever experimented at the pageload level. We experimented at the user level, so 1% of users would get the new UI. I suppose this is only possible at the millions of users scale which all of them had.
Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?
21–30 of 817 posts
Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?
#22Yes. Before the update, when its avatar was still black, it solved pretty complex coding problems effortlessly and gave very nuanced, thoughtful answers to non-programming questions. Now it struggles with just changing two lines in a 10-line block of CSS and printing this modified 10-line block again. Some lines are missing, others are completely different for no reason. I'm sure scaling the model is hard, but they l…
Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?
#23Earlier quoted context omitted.
No place I worked at ever experimented at the pageload level. We experimented at the user level, so 1% of users would get the new UI. I suppose this is only possible at the millions of users scale which all of them had.
I updated the comment to reflect that. Certainly the signal is stronger because you’re amortizing away the surprise factor of the change, and at least it’s a consistent UX, but the UX tradeoff in the worst case is that experiment-group users get a broken product with no notice or escape hatch. Unless you’re being very careful, meticulous, and transparent it’s just not acceptable if you’re a paying customer.
Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?
#24To me, it feels like it's started giving superficial responses and encouraging follow-up elsewhere -- I wouldn't be surprized if its prompt has changed to something to that effect. Before, if I had an issue with a library or debugging issue, it would try to be helpful and walk me through potential issues, and ask me to 'let it know' if it worked or not. Now it will try to superficially diagnose the problem and then a…
Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?
#25The answer is the same on GPT plus and API with GPT-4, even with "developer" role.
Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?
#26For a while, if you asked the iPhone version what it was it claimed to be GPT3.0. Not sure if it still is that, but I noticed the iPhone version was a bit worse. Maybe they rolled that out more broadly?
Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?
#27Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?
#28Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?
#29It’s possible they are trying out a shaved/turbo version so that they can start removing the limits. I mean as it is - 25 messages every 3 hours is useless, particularly for browsing and plugins.
Imagine trading the advice of a senior mentor for 5 intermediate mentors. Yes, the answers get to you faster, but it's much less useful.