Live data from Hacker News

Experiencing decreased performance with ChatGPT-4

community.openai.com

31–40 of 200 posts

Re: Experiencing decreased performance with ChatGPT-4

#31

I’m convinced this is group hallucination. It must be so interesting to work at OpenAI, knowing you didn’t change a thing, and seeing that because of random chance, some small fraction of 100M users have all tricked each other that suddenly, something is different.

But they do change things in the ChatGPT web app. You can't choose the exact model there, just 3/4 and from time to time they update the models they use.

Re: Experiencing decreased performance with ChatGPT-4

#32
I feel like this gets posted at least once a week. I even posted about v3 a while back.

My gut feeling, based on no evidence, is that with the constant pruning of whatever base prompt they're using to seed conversations, the overly-strict rules by which the model is allowed to generate responses is causing it to have worse and worse outputs.

It really started getting bad when ClosedAI began to add all of the policies about what's "allowed", e.g. that it's not allowed to generate silly but non-factual information.

Re: Experiencing decreased performance with ChatGPT-4

#33

I’m convinced this is group hallucination. It must be so interesting to work at OpenAI, knowing you didn’t change a thing, and seeing that because of random chance, some small fraction of 100M users have all tricked each other that suddenly, something is different.

How do you make something use substantially less resources without changing a thing?

Re: Experiencing decreased performance with ChatGPT-4

#34
I wonder if the pre-prompting part was increased as part of the trust/safety effort. For example they increased more examples of the types of things not to say, attached to the user prompts. That would decrease the amount of reasoning it could give the actual prompt, as it has to logically make sure each statement complies with the prior rules, and it also decreases the context length.

Re: Experiencing decreased performance with ChatGPT-4

#35

I don't know about the API, but the ChatGPT UI GPT4 really became much worse, so much so that I had to cancel my subscription. It's not just about novelty factor. I used to store all my prompts locally, and when I compare old responses and new ones, there is a huge difference. OpenAI employee said the API model doesn't change, but were careful to not say anything about the UI. I am now waiting to get access to the AP…

Can you give an example?

Re: Experiencing decreased performance with ChatGPT-4

#37
post #28

The thing is they(OpenAI) really could mess with the temperature/probability settings, and other settings for various reasons that can cause something like this to happen to some people at certain random times. Hence is hard to reproduce, and only OpenAI knows it, either they acknowledge or not that is the based on business requirements

Right and they can change those settings and say that it's the same model and they would be technically telling the precise truth.

Re: Experiencing decreased performance with ChatGPT-4

#38
I use the following test to ensure I'm on GPT4 and not 3.5. (I noticed that it did fail at this test temporarily and then got it. Not sure why. Maybe it reverts back to 3.5 when under load?)

I have a 12 liter jug and a 6 liter jug. I want to measure 6 liters. How do I do it?

GPT4: You actually don't need to do anything because one of your jugs is already a 6-liter jug. If you fill it up to the top, you'll have exactly 6 liters of water.

GPT-3.5: To measure exactly 6 liters using a 12-liter jug and a 6-liter jug, you can follow the steps below:

Start with both jugs empty. Fill the 12-liter jug completely with water. Pour the water from the 12-liter jug into the 6-liter jug. This will leave you with 6 liters of water in the 12-liter jug. Empty the 6-liter jug. Pour the 6 liters of water from the 12-liter jug back into the empty 6-liter jug. Now, you have 6 liters of water in the 6-liter jug. At this point, you have successfully measured 6 liters using the 12-liter jug and the 6-liter jug.

Re: Experiencing decreased performance with ChatGPT-4

#39
Notice how you never hear anyone saying that GPT-4 is better since the launch. You'd expect to hear something like that as people gain more experience with prompting it.

I've certainly noticed that the quality of responses has gone down, and I have to repeat myself more often as it doesn't always remember all my instructions.

For an example of something it can no longer do, I used to show it off by having it explain something using words that each start with the next letter of the alphabet, then I'd add "now make it rhyme" after it succeeded. If you try that now (even with the 0314 model), it'll fail at the task.

Re: Experiencing decreased performance with ChatGPT-4

#40

I don't know about the API, but the ChatGPT UI GPT4 really became much worse, so much so that I had to cancel my subscription. It's not just about novelty factor. I used to store all my prompts locally, and when I compare old responses and new ones, there is a huge difference. OpenAI employee said the API model doesn't change, but were careful to not say anything about the UI. I am now waiting to get access to the AP…

why will nobody provide these prompts holy shit this is such a frustrating topic
Post reply on HN