LLaMA2 Chat 70B outperformed ChatGPT
tatsu-lab.github.io
LLaMA2 Chat 70B outperformed ChatGPT
1–10 of 135 posts
Re: LLaMA2 Chat 70B outperformed ChatGPT
#2Re: LLaMA2 Chat 70B outperformed ChatGPT
#3Re: LLaMA2 Chat 70B outperformed ChatGPT
#4*edit: oops, my brain inserted "by" in the middle of "outperformed chatgpt". I'll leave my wrong comment up as a testament to shame.
Re: LLaMA2 Chat 70B outperformed ChatGPT
#5Disclaimer from the site:
> Caution: GPT-4 may favor models with longer outputs and/or those that were fine-tuned on GPT-4 outputs.
> While AlpacaEval provides a useful comparison of model capabilities in following instructions, it is not a comprehensive or gold-standard evaluation of model abilities. For one, as detailed in the AlpacaFarm paper, the auto annotator winrates are correlated with length.
Re: LLaMA2 Chat 70B outperformed ChatGPT
#6Re: LLaMA2 Chat 70B outperformed ChatGPT
#7I haven't had a chance to use the GPT-4 API yet - is it that much better than the GPT-4 available via ChatGPT? Or am I misunderstanding?
Re: LLaMA2 Chat 70B outperformed ChatGPT
#8Re: LLaMA2 Chat 70B outperformed ChatGPT
#9* ChatGPT 3.5. But it's also within spitting distance of GPT4, which is very exciting.
Re: LLaMA2 Chat 70B outperformed ChatGPT
#10I haven't had a chance to use the GPT-4 API yet - is it that much better than the GPT-4 available via ChatGPT? Or am I misunderstanding?