Live data from Hacker News

Alpaca RLHF-ed to beat ChatGPT

crfm.stanford.edu

41–45 of 45 posts

Re: Alpaca RLHF-ed to beat ChatGPT

#41

Whenever I see a claim about GPT I get temporarily interested until I learn it’s GPT3.5 and not GPT4. 4 isn’t just marginally better at most tasks I use it for, it’s operating at an entirely different level to the point where I have little (no?) day-to-day use of 3.5 at this point.

I'm assuming you use gpt4 via ChatGPT plus. Does the message cap bother you? I heard it's something like 25 messages per 3 hours. That sounds so low I don't even bother subscribing. I guess this doesn't apply if you use it via the api.

I have no idea what happened. My original 1-month Plus subscription to test GPT-4 was, as usual, limited to 25/hour and tied to my personal gmail address. But then recently-- in the last week-- I resubscribed and I accidentally did so under my work email, which has a more specialized and restricted TLD, and under that I don't have any quota limits for the GPT-4 version of ChatGPT.

It might seem a small thing, having to space out prompts 25 ever 3 hours when you might not have used more than 100-200 in a day anyway, but the net result is liberating. I experiment, explore the limits, and get whimsical with it to a much greater extent than when I can to consciously think about each prompt as a rationed resource.

Re: Alpaca RLHF-ed to beat ChatGPT

#42

Whenever I see a claim about GPT I get temporarily interested until I learn it’s GPT3.5 and not GPT4. 4 isn’t just marginally better at most tasks I use it for, it’s operating at an entirely different level to the point where I have little (no?) day-to-day use of 3.5 at this point.

I'm assuming you use gpt4 via ChatGPT plus. Does the message cap bother you? I heard it's something like 25 messages per 3 hours. That sounds so low I don't even bother subscribing. I guess this doesn't apply if you use it via the api.

For me tbh, i kinda like the limit. I use GPT-4 a lot lately. Hitting the limit reminds me, i got too lazy writing code myself or i got way too deep into it. Then i just close the tab & remember, that i still love writing code the (not quite yet) old way.

Re: Alpaca RLHF-ed to beat ChatGPT

#43
post #6

Absolutely off topic, but I just got back from a week in Peru, where the alpaca is a prominent member of the local fauna. For folks in the US at least, it's a relatively inexpensive trip and an absolutely gobsmackingly gorgeous country with friendly people and amazing food. Highly recommended!!!

Interesting, did other local fauna converse with you in standard 5-paragraph essay formats as taught to humans >= 12 years old, or was it only the alpaca that did so?

Re: Alpaca RLHF-ed to beat ChatGPT

#44
Beating by generating longer answer is not a win for me. Maybe raters prefer long answers, but in reality long answers are only good if they provide extra important information.

They should try to compare answers with similar length.

Re: Alpaca RLHF-ed to beat ChatGPT

#45

Earlier quoted context omitted.

I wonder how long "Using humans to rate the quality of other humans" thing can last. Surely academia has only so long before it collapses.

You're asserting that current LLMs are as capable as evaluating each other as are humans with advanced degrees?

Yes. They're both awful.
Post reply on HN