Live data from Hacker News

ShareGPT: Share your ChatGPT conversations with one click

sharegpt.com

41–42 of 42 posts

Re: ShareGPT: Share your ChatGPT conversations with one click

#41
post #39
post #16

Earlier quoted context omitted.

Because he has people skills.

They call it prompt engineer nowadays. Some are experts at engineering prompts to human resources, other engineer prompts to artificial resources.

So literally a job which exists to only be replaced by the thing it's feeding. I guess the memes back in the early 00's of Google being a data octopus will evolve into the OpenAI Octopus eating those very humans.

Re: ShareGPT: Share your ChatGPT conversations with one click

#42

Earlier quoted context omitted.

In chess there is a very clear victory state, and a scoring function can be implicitly defined from a large number of games of various skill levels. You really don't throw two sentences into the thunder dome to decide which one "wins". Means it's much more susceptible to being poisoned.

>You really don't throw two sentences into the thunder dome to decide which one "wins". That's almost literally what RLHF is though, and that is the last step of training GPT-n. Then when GPT-{n+1} is being trained, it will include some results from GPT-n, and therefore will benefit from that finetuning, even before it goes through its own round of RLHF. Also, on average good outputs of GPT-n are more likely to be in…

I suspect the comment about the thunder dome was a reference to RLHF. On the one hand RLHF seems far superior to the kind of prompt engineering Microsoft seems to have relied on with Sydney. On the other, it's dubious that the manual selection in RLHF is really always selecting for quality, as against at least to some significant extent pandering to whatever biases or preferences the humans in the training loop might have.
Post reply on HN