Live data from Hacker News

Open-source reflection lama 70B beats Claude 3.5 and GPT-4 on benchmarks

reflectionllama.com

1–6 of 6 posts

Re: Open-source reflection lama 70B beats Claude 3.5 and GPT-4 on benchmarks

#4

Related recent discussion: Reflection 70B, the top open-source model https://news.ycombinator.com/item?id=41459781

That's a great conversation about the reflection 70b, but do you still have doubts about whether it's a hype or a game-changer?

The link provided leads to a playground for the reflection llama 70B.

Re: Open-source reflection lama 70B beats Claude 3.5 and GPT-4 on benchmarks

#6
post #4

Related recent discussion: Reflection 70B, the top open-source model https://news.ycombinator.com/item?id=41459781

That's a great conversation about the reflection 70b, but do you still have doubts about whether it's a hype or a game-changer? The link provided leads to a playground for the reflection llama 70B.

In the Physics of Language Models talk[1], he shows how a LLM trained to be able to backtrack can give much better answers.

However he also points out it has to be included in the initial training, you can't improve a non-backtrack-trained model by finetuning it later.

So seems it's probably the way to go for training new models, but limited applicability to those already trained.

[1]: https://www.youtube.com/watch?v=yBL7J0kgldU