Live data from Hacker News

Code Llama, a state-of-the-art large language model for coding

ai.meta.com

261–270 of 525 posts

Re: Code Llama, a state-of-the-art large language model for coding

#261

no more work soon?

The ability to work less historically has always came as a byproduct of individuals earning more per hour through productivity increases.

The end goal of AI isn't to make your labour more productive, but to not need your labour at all.

As your labour becomes less useful if anything you'll find you need to work more. At some point you may be as useful to the labour market as someone with 60 IQ today. At this point most of the world will become entirely financially dependent on the wealth redistribution of the few who own the AI companies producing all the wealth – assuming they take pity on you or there's something governments can actually do to force them to pay 90%+ tax rates, of course.

Re: Code Llama, a state-of-the-art large language model for coding

#262

Earlier quoted context omitted.

That’s interesting. I tend to lump FB, Amazon, Google, and MS in my head when thinking about the tech giants, but you’re right, FB is the only one of those not offering a commercial platform. For them, building out the capabilities of the LLMs is something to be done in the open with community involvement, because they’re not going to monetize the models themselves. They’re also getting a fantastic amount of press fr…

FB is unlike the other BigTech(tm) since Zuck never sold out and has a controlling equity stake. Amazon, Google, and MS are all controlled by and beholden to institutional investors. FB can release these for no other reason than Zuck’s ego or desire to kill OpenAI. Same deal as him going off on a tangent with the Metaverse thing.

Larry and Sergey still control majority of voting power from what I recall.

Re: Code Llama, a state-of-the-art large language model for coding

#264

TheBloke doesn’t joke around [1]. I’m guessing we’ll have the quantized ones by the end of the day. I’m super excited to use the 34B Python 4 bit quantized one that should just fit on a 3090. [1] https://huggingface.co/TheBloke/CodeLlama-13B-Python-fp16

What kind of cpu/gpu power do you need for quantization or these new gguf formats ?

i run llama2 13B models with 4-6 k-quantized oin a 3060 with 12Gb VRam

Re: Code Llama, a state-of-the-art large language model for coding

#265
post #242

Earlier quoted context omitted.

That’s interesting. I tend to lump FB, Amazon, Google, and MS in my head when thinking about the tech giants, but you’re right, FB is the only one of those not offering a commercial platform. For them, building out the capabilities of the LLMs is something to be done in the open with community involvement, because they’re not going to monetize the models themselves. They’re also getting a fantastic amount of press fr…

I'm in the camp that its a mistake for Meta to not be competing in the commercial compute space. wrote about and diagramed it here - https://telegra.ph/Facebook-Social-Services-FbSS-a-missed-op...

Meta absolutely could not overcome the barriers to entry and technical mismatch for any sort of traditional IAS style product, and it would be foolish for them to try. They might be able to pull off some sort of next generation Heroku style service aimed at smaller shops with built in facebook integration and authn/z management, but that's tangential.

Re: Code Llama, a state-of-the-art large language model for coding

#266

It's really sad how everyone here is fawning over tech that will destroy you own livelihoods. "AI won't take your job, those who use AI will" is purely short term, myopic thinking. These tools are not aimed to help workers, the end goal is to make it so you don't need to be an engineer to build software, just let the project manager or director describe the system they want and boom there it is. You can scream that t…

"If software engineering becomes a solved problem" It will simply move to a higher level of abstraction. Remind me, how many programmers today are writing in assembly?

due to the way LLMs work: it will be able to handle that level of abstraction in exactly the same way

Re: Code Llama, a state-of-the-art large language model for coding

#267

How are people using these local code models? I would much prefer using these in-context in an editor, but most of them seem to be deployed just in an instruction context. There's a lot of value to not having to context switch, or have a conversation. I see the GitHub copilot extensions gets a new release one every few days, so is it just that the way they're integrated is more complicated so not worth the effort?

http://cursor.sh integrates GPT-4 into vscode in a sensible way. Just swapping this in place of GPT-4 would likely work perfectly. Has anyone cloned the OpenAI HTTP API yet?

LocalAI https://localai.io/ and LMStudio https://lmstudio.ai/ both have fairly complete OpenAI compatibility layers. llama-cpp-python has a FastAPI server as well: https://github.com/abetlen/llama-cpp-python/blob/main/llama_... (as of this moment it hasn't merged GGUF update yet though)

Re: Code Llama, a state-of-the-art large language model for coding

#268
post #7

Does anyone have a good explanation for Meta's strategy with AI? The only thing I've been able to think is they're trying to commoditize this new category before Microsoft and Google can lock it in, but where to from there? Is it just to block the others from a new revenue source, or do they have a longer game they're playing?

It makes sense to me that Facebook is releasing these models similarly to the way that Google releases Android OS. Google's advertising model benefits from as many people being online as possible and their mobile operating system furthers that aim. Similarly, Facebook's advertising model benefits from having loads of content being generated to then be posted in their various products' feeds.

Re: Code Llama, a state-of-the-art large language model for coding

#269
post #7

Does anyone have a good explanation for Meta's strategy with AI? The only thing I've been able to think is they're trying to commoditize this new category before Microsoft and Google can lock it in, but where to from there? Is it just to block the others from a new revenue source, or do they have a longer game they're playing?

Clearly the research team at Meta knows the domain as well anybody, has access to a data trove as large as anybody and their distribution capability is as large scale as anyone's. If their choice right now is not to try to overtly monetize these capabilities but instead commoditize and "democratize" what others are offering it suggests they think that a proprietary monetization route is not available to them. In othe…

There really is no technical moat. Any new architectures are going to be published because that's 100% the culture and AI folks won't work somewhere where that's not true. Training details/hyperparameters/model "build-ons" aren't published but those are a very weak moat.

The only moat that is meaningful is data and they've got that more than any other player save maybe google. Publishing models doesn't erode that moat, and it's not going anywhere as long as facebook/whatsapp/instagram rule "interactive" social.

Re: Code Llama, a state-of-the-art large language model for coding

#270
post #122

Never before in the history of mankind was a group so absolutely besotted with the idea of putting themselves out of a job.

I understand the fear of losing your job or becoming less relevant, but many of us love this work because we're passionate about technology, programming, science, and the whole world of possibilities that this makes... possible.

That's why we're so excited to see these extraordinary advances that I personally didn't think I'd see in my lifetime.

The fear is legitimate and I respect the opinions of those who oppose these advances because they have children to provide for and have worked a lifetime to get where they are. But at least in my case, the curiosity and excitement to see what will happen is far greater than my little personal garden. Damn, we are living what we used to read in the most entertaining sci-fi literature!

(And that's not to say that I don't see the risks in all of this... in fact, I think there will be consequences far more serious than just "losing a job," but I could be wrong)

Post reply on HN