Code Llama, a state-of-the-art large language model for coding
41–50 of 525 posts
Re: Code Llama, a state-of-the-art large language model for coding
#42Re: Code Llama, a state-of-the-art large language model for coding
#43Does anyone have a good explanation for Meta's strategy with AI? The only thing I've been able to think is they're trying to commoditize this new category before Microsoft and Google can lock it in, but where to from there? Is it just to block the others from a new revenue source, or do they have a longer game they're playing?
If you watch the Connect talks, I'll be speaking about this..
Re: Code Llama, a state-of-the-art large language model for coding
#44Earlier quoted context omitted.
>Even the 7B model of code llama seems to be competitive with Codex, the model behind copilot It's extremely good. I keep a terminal tab open with 7b running for all of my "how do I do this random thing" questions while coding. It's pretty much replaced Google/SO for me.
You've already downloaded and thoroughly tested the 7B parameter model of "code llama"? I'm skeptical.
Re: Code Llama, a state-of-the-art large language model for coding
#45The highlight IMO > The Code Llama models provide stable generations with up to 100,000 tokens of context. All models are trained on sequences of 16,000 tokens and show improvements on inputs with up to 100,000 tokens. Edit: Reading the paper, key retrieval accuracy really deteriorates after 16k tokens, so it remains to be seen how useful the 100k context is.
Re: Code Llama, a state-of-the-art large language model for coding
#46Earlier quoted context omitted.
You've already downloaded and thoroughly tested the 7B parameter model of "code llama"? I'm skeptical.
Just sign up at meta and you'll get an email link in like 5 minutes
No one who has been using any model for just the past 30 minutes would say that it has "pretty much replaced Google/SO" for them, unless they were being facetious.
Re: Code Llama, a state-of-the-art large language model for coding
#47Re: Code Llama, a state-of-the-art large language model for coding
#48No thanks, going back to Winamp.
Re: Code Llama, a state-of-the-art large language model for coding
#49Does anyone have a good explanation for Meta's strategy with AI? The only thing I've been able to think is they're trying to commoditize this new category before Microsoft and Google can lock it in, but where to from there? Is it just to block the others from a new revenue source, or do they have a longer game they're playing?
because language models are a complementary product, and the complement must be commoditized as a strategy.
I see AMD as a bigger beneficiary, since, very soon, amd will equal nvidia for inference and fine-tuning, but amd has a long way to go to equal in foundation model training.