Live data from Hacker News

Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?

news.ycombinator.com

111–120 of 817 posts

Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?

#111
post #55
post #54

Earlier quoted context omitted.

Genuinely asking, what's an "unwoke" prompt?

Something that follows my actual requests, without trying to lecture me about feminism and other U.S. Democrats topics.

What is an example of a request that is causing these issues?

Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?

#112

Earlier quoted context omitted.

Try out Bard, it's coding is much improved in the last 2 weeks. I've unfortunately switched over for the time being.

No thanks! I have better things to do than feeding that advertising behemoth. What I like about ChatGPT is that I don't see any ads at all!

> What I like about ChatGPT is that I don't see any ads at all!

For now. It's just a marketing tool/demo site, like ITA Matrix was/is. The ads are vended by Bing.

Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?

#113
post #38

Earlier quoted context omitted.

I’d like to see a model with the effluent of the internet intelligently filtered from the pretraining data by LLM and human curation, and much more effort to include digitised archival sources and the entirety of books and high quality media transcripts. I imagine it would yield far better baseline quality outputs with much less than current “requirements” for (over)correction with ultimately disastrous RLHF masking.

I'd love to play with a version of GPT 4 fine-tuned with every science textbook written in the last few decades, every published science paper (not just preprints from ArXiV), and everything generated by every large research institute. Think NASA, CERN, etc... Or one tuned with every fiction novel ever written, along with every screenplay.

So a model fine-tuned on libgen?

Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?

#114

Earlier quoted context omitted.

It was a great ride while it lasted. My assumption is that efficacy at coding tasks is such a small percent of users, they’ve just sacrificed it on the altar of efficiency and/or scale. That, or they’ve cut some back room deal with Microsoft to make Copilot have access to the only version of the model that can actually code.

FWIW, I started to get the same feeling as the OP about GPT-4 model I have access to on Azure, so if there's any deal being cut here, it might involve dumbing down the model for paying Azure customers as well. Now, to be clear: I only started to get a feeling that GPT-4 on Azure is getting worse. I didn't do any specific testing for this so far, as I thought I may just be imagining it. This thread is starting to conv…

I’ve seen degradation in the app and via the API, so if I had to bet, they’ve probably kneecapped the model so that it works passably everywhere they’ve been made it available vs. works well in one place or another.

Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?

#116
Yes, I noticed this too, I fed it some HTML with tailwind classes and told it to just list all the tailwind classes that we use and then the CSS behind those classes.. it just hallucinated all(!) the items in the list (and just gave me a list of 10 seemingly random classes). And then when I did asked something else about the code it had forgotten I had ever pasted anything in the conversation. Very weird.

Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?

#117
post #14

Yes. Before the update, when its avatar was still black, it solved pretty complex coding problems effortlessly and gave very nuanced, thoughtful answers to non-programming questions. Now it struggles with just changing two lines in a 10-line block of CSS and printing this modified 10-line block again. Some lines are missing, others are completely different for no reason. I'm sure scaling the model is hard, but they l…

Try out Bard, it's coding is much improved in the last 2 weeks. I've unfortunately switched over for the time being.

Google (Deepmind) actually has the people and has developed the science to make the best AI products in the world, but unfortunately Bard seems to be thrown together in an afternoon by an intern, and then handed off to a hoard of marketing people. It's not good right now. Deepmind is one of the best scientifically, they just don't really make products. OpenAI is essentially the direct opposite of that.

Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?

#118
post #89

Earlier quoted context omitted.

Not necessarily American, you just have to avoid EU and, I believe, Russia/China/Cuba etc.

I'm in Switzerland and Bard is locked out, we do not go by EU laws because we are not part of the EU. We have plenty of bilateral deals but still.

But don't you sill have privacy laws very similar to the GDPR?

Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?

#119
post #14

Yes. Before the update, when its avatar was still black, it solved pretty complex coding problems effortlessly and gave very nuanced, thoughtful answers to non-programming questions. Now it struggles with just changing two lines in a 10-line block of CSS and printing this modified 10-line block again. Some lines are missing, others are completely different for no reason. I'm sure scaling the model is hard, but they l…

If this is true, one should be able to compare with benchmarks or evals to demonstrate this.

Anyone know more about this?

Re: Ask HN: Is it just me or GPT-4's quality has significantly deteriorated lately?

#120

Naturally, people feed it shit in order to bend results for others for their gain. Everyone who was in crypto manipulations, now rushed in this field, and they are extremely smart and extremely ruthless people, and they have no legal limits being anonymous, and almost unlimited funding they made in crypto. It reduces results quality, and also invokes countermeasures on their side to limit damage, that further reduces…

In general data from conversations isn't merely instantly fed back into the model so there is no way for for users to feed garbage into the model in the fashion you imagine.
Post reply on HN