4 isn’t just marginally better at most tasks I use it for, it’s operating at an entirely different level to the point where I have little (no?) day-to-day use of 3.5 at this point.
Alpaca RLHF-ed to beat ChatGPT
21–30 of 45 posts
Re: Alpaca RLHF-ed to beat ChatGPT
#22Hm. Title: “beats Chat GPT” Reality: > With these evaluation instructions, we compare RLHF model responses to Davinci003 responses and measure the fraction of times the RLHF model is preferred; we call this statistic the win-rate. > Of the methods we studied, PPO proves the most effective, improving the win-rate against Davinci003 from 44% to 55% according to human evaluation, which even outperforms ChatGPT. …for the…
Re: Alpaca RLHF-ed to beat ChatGPT
#23Whenever I see a claim about GPT I get temporarily interested until I learn it’s GPT3.5 and not GPT4. 4 isn’t just marginally better at most tasks I use it for, it’s operating at an entirely different level to the point where I have little (no?) day-to-day use of 3.5 at this point.
Re: Alpaca RLHF-ed to beat ChatGPT
#24Whenever I see a claim about GPT I get temporarily interested until I learn it’s GPT3.5 and not GPT4. 4 isn’t just marginally better at most tasks I use it for, it’s operating at an entirely different level to the point where I have little (no?) day-to-day use of 3.5 at this point.
I guess this doesn't apply if you use it via the api.
Re: Alpaca RLHF-ed to beat ChatGPT
#25Whenever I see a claim about GPT I get temporarily interested until I learn it’s GPT3.5 and not GPT4. 4 isn’t just marginally better at most tasks I use it for, it’s operating at an entirely different level to the point where I have little (no?) day-to-day use of 3.5 at this point.
I'm assuming you use gpt4 via ChatGPT plus. Does the message cap bother you? I heard it's something like 25 messages per 3 hours. That sounds so low I don't even bother subscribing. I guess this doesn't apply if you use it via the api.
Re: Alpaca RLHF-ed to beat ChatGPT
#26Whenever I see a claim about GPT I get temporarily interested until I learn it’s GPT3.5 and not GPT4. 4 isn’t just marginally better at most tasks I use it for, it’s operating at an entirely different level to the point where I have little (no?) day-to-day use of 3.5 at this point.
I'm assuming you use gpt4 via ChatGPT plus. Does the message cap bother you? I heard it's something like 25 messages per 3 hours. That sounds so low I don't even bother subscribing. I guess this doesn't apply if you use it via the api.
When I hit the limit, I work on the problem myself and wait until 4 resets instead of relying on 3.5. 4 is so much better that I don’t trust 3.5 with my work anymore.
Re: Alpaca RLHF-ed to beat ChatGPT
#27Whenever I see a claim about GPT I get temporarily interested until I learn it’s GPT3.5 and not GPT4. 4 isn’t just marginally better at most tasks I use it for, it’s operating at an entirely different level to the point where I have little (no?) day-to-day use of 3.5 at this point.
I'm assuming you use gpt4 via ChatGPT plus. Does the message cap bother you? I heard it's something like 25 messages per 3 hours. That sounds so low I don't even bother subscribing. I guess this doesn't apply if you use it via the api.
Your coding speed is unlikely to be that fast, requiring 25 code segments in 3 hours. GPT-4 outputs something, you need time to double check, test, additional googling etc. Its still a massive speed boost.
Using it recreationally (Especially chatting) will result in a lot more requests.
Re: Alpaca RLHF-ed to beat ChatGPT
#28Whenever I see a claim about GPT I get temporarily interested until I learn it’s GPT3.5 and not GPT4. 4 isn’t just marginally better at most tasks I use it for, it’s operating at an entirely different level to the point where I have little (no?) day-to-day use of 3.5 at this point.
I'm assuming you use gpt4 via ChatGPT plus. Does the message cap bother you? I heard it's something like 25 messages per 3 hours. That sounds so low I don't even bother subscribing. I guess this doesn't apply if you use it via the api.
I haven't used it a whole lot.
Re: Alpaca RLHF-ed to beat ChatGPT
#29Earlier quoted context omitted.
I'm assuming you use gpt4 via ChatGPT plus. Does the message cap bother you? I heard it's something like 25 messages per 3 hours. That sounds so low I don't even bother subscribing. I guess this doesn't apply if you use it via the api.
I was initially deterred. But in practice, when using it for professional purposes, I never encounter it. Your coding speed is unlikely to be that fast, requiring 25 code segments in 3 hours. GPT-4 outputs something, you need time to double check, test, additional googling etc. Its still a massive speed boost. Using it recreationally (Especially chatting) will result in a lot more requests.
Re: Alpaca RLHF-ed to beat ChatGPT
#30Earlier quoted context omitted.
I wonder how long "Using humans to rate the quality of other humans" thing can last. Surely academia has only so long before it collapses.
You're asserting that current LLMs are as capable as evaluating each other as are humans with advanced degrees?