There's some blatant astroturfing from new accounts going on in this thread - I gotta say it's not the best impression
Free Dolly: First truly open instruction-tuned LLM
31–40 of 70 posts
Re: Free Dolly: First truly open instruction-tuned LLM
#32Re: Free Dolly: First truly open instruction-tuned LLM
#33Re: Free Dolly: First truly open instruction-tuned LLM
#34Any info on Pythia base model performance versus GPT-3 or 3.5? Couldn't find any benchmarks in the paper. I imagine LLaMA is ahead there.
Re: Free Dolly: First truly open instruction-tuned LLM
#35There's some blatant astroturfing from new accounts going on in this thread - I gotta say it's not the best impression
Re: Free Dolly: First truly open instruction-tuned LLM
#36Why is this post full of n00b user 'comments'?
Re: Free Dolly: First truly open instruction-tuned LLM
#37Why is this post full of n00b user 'comments'?
@dang, any chance we can just ban all these accounts? Seems to be pretty cut and dry here.
Re: Free Dolly: First truly open instruction-tuned LLM
#38There's some blatant astroturfing from new accounts going on in this thread - I gotta say it's not the best impression
Indeed, I'm not sure how to summon @dang but the number of databricks shill comments in this thread is absurd.
Re: Free Dolly: First truly open instruction-tuned LLM
#39Re: Free Dolly: First truly open instruction-tuned LLM
#40Any info on Pythia base model performance versus GPT-3 or 3.5? Couldn't find any benchmarks in the paper. I imagine LLaMA is ahead there.
Do we have any quantitative way of benchmarking the quality of these models at all? Like, I don’t care if a model takes one minute per token on my laptop as long as it’s “GPT-4 quality”, and I don’t care if it does 100 tokens per second if it’s straight crap. But every comparison I see people make regarding quality seems to come from “I asked it a couple of my favorite questions and it did… uh, only a little worse th…