Live data from Hacker News

IBM Granite: A Family of Open Foundation Models for Code Intelligence

github.com

41–50 of 78 posts

Re: IBM Granite: A Family of Open Foundation Models for Code Intelligence

#41

I wonder why companies like IBM are jumping on the LLM bandwagon and training/releasing models that have no chance of competing with Llama/Mistral? To me it just looks like a complete waste of $$ because nobody will use them in any serious scenarios

Same reason they jumped on the clown bandwagon, it's the kind of offering it's expected to have when you're a company like that. Huge size, leading research departments, big enterprise customers.

They've been doing "AI" for ages. Notably Watson over the last couple of decades or so.

Re: IBM Granite: A Family of Open Foundation Models for Code Intelligence

#42
post #7

Earlier quoted context omitted.

IBM made $60 billion in revenue last year. Where do you think it all came from? The same companies/governments that buy their overpriced crap are going to buy these new LLMs as well.

These are open weight models released under an Apache 2.0 license. There's nothing to buy.

IBM is a sales and services org.

Their customers aren’t going to build their own RAG and agent frameworks, vector DBs, data ingest pipelines, finetunes, high scale inference serving solutions, etc, etc.

There’s an incredible amount of stuff to buy.

Re: IBM Granite: A Family of Open Foundation Models for Code Intelligence

#43
post #7

Earlier quoted context omitted.

IBM made $60 billion in revenue last year. Where do you think it all came from? The same companies/governments that buy their overpriced crap are going to buy these new LLMs as well.

These are open weight models released under an Apache 2.0 license. There's nothing to buy.

Who is going to host them?

Re: IBM Granite: A Family of Open Foundation Models for Code Intelligence

#44

Earlier quoted context omitted.

IBM goes at great lengths to train models on clean data that has lower risk of copyright or legal issues attached. Just take a look at the model description. That data issue is important enough for some companies to pick mediocre model over llama or mistral.

What if I told you that a lot of freely licensed code on GitHub is not clean? That the authors may have read something and rewritten it in a way that wasn’t transformative? So it basically has the same problems.

What if I told you the supposedly clean "The Stack" dataset contains at least one GPL repository inside, just because their license detection tool bugged out?

IBM and other big players are vigilant about these things, and this is what companies pay for.

Their software may not be better in some metrics, but they're cleaner in some and their support contracts allows people to sleep tight at night.

This is what money buys. Peace of mind and continuity.

Re: IBM Granite: A Family of Open Foundation Models for Code Intelligence

#45
post #12

Earlier quoted context omitted.

IBM do a mixture of shovelware and extremely hardcore tech so they could honestly go either way with this.

> mixture of shovelware and extremely hardcore tech Citation needed All I've seen from them in my professional experience is actually legacy mainframe maintenance.. Not shovelware, but very far from hardcore tech.

On the contrary, the maintenance and continued improvement of an entire ISA and ISA specific operating systems is exactly my idea of hardcore tech, i.e. continuing to pay a chip org to design new chips for said ISA every generation and implement new instructions...and continuing to pay OS and compiler programmers to work those into their OS's and compilers...I'm not sure where we draw the line on maintenance vs. continued development here, but I'm not sure I'd call that purely maintenence.

There really aren't a lot of companies out there that can claim to do similar (and of course besides s390x, an ancient and venerable CISC, IBM also has Power, so they are doing this 2x over). You'll find a lot of IBM employees contributing to what I'd consider "hardcore" tech like LLVM and the Linux kernel as a result, because they genuinely have a large amount of expertise in those and similar areas. And here I'm not even really including Red Hat, but if you include them then they are even more overweight in the hardcore tech category.

If anything, a lot of the rest of the tech industry has left "hardcore tech" behind due to efficiency concerns as a result of a longrunning industry wide process of consolidation and commodification that IBM has resisted for obvious reasons. IBM is hardcore to a fault if anything.

TLDR: I actually think IBM punches above their weight in the "hardcore tech" area so long as our definition is sufficiently low level rather than say, cloud services, in which case fair enough you can probably fairly say they suck at that.

Here I've also chosen to entirely ignore IBM research.

Re: IBM Granite: A Family of Open Foundation Models for Code Intelligence

#46
post #37
post #16

Earlier quoted context omitted.

Llama and Mistral are already local & fulfill these requirements

Can I sue Lamma and Mistral if things go wrong?

Llama is owned by Meta, so you’d be suing meta

But I’m pretty sure both models have “we’re not responsible” clauses.

Re: IBM Granite: A Family of Open Foundation Models for Code Intelligence

#47
post #33

Earlier quoted context omitted.

By no means it is an outdated phrase. Ask any startup sales person!

I personally know a VP who was fired for buying "IBM Cloud." You can absolutely get flak for choosing IBM these days, even at a stodgy enterprise. The gist is still current, but you need to fill in AWS as the current uncontroversial choice.

Must be a very terrible company to work for if they are firing people solely on them picking X over Y.

Re: IBM Granite: A Family of Open Foundation Models for Code Intelligence

#48

Earlier quoted context omitted.

> mixture of shovelware and extremely hardcore tech Citation needed All I've seen from them in my professional experience is actually legacy mainframe maintenance.. Not shovelware, but very far from hardcore tech.

On the contrary, the maintenance and continued improvement of an entire ISA and ISA specific operating systems is exactly my idea of hardcore tech, i.e. continuing to pay a chip org to design new chips for said ISA every generation and implement new instructions...and continuing to pay OS and compiler programmers to work those into their OS's and compilers...I'm not sure where we draw the line on maintenance vs. cont…

[deleted]

Re: IBM Granite: A Family of Open Foundation Models for Code Intelligence

#49

https://i.kym-cdn.com/photos/images/original/001/138/631/b7a...

Is this a segway for IBM to release Terraform specific LLMs so I never have to write that hot garbage ever again? Sign me up IBM!

Here's a similar existing product- https://www.ibm.com/products/watsonx-code-assistant-ansible-...

Re: IBM Granite: A Family of Open Foundation Models for Code Intelligence

#50

I wonder why companies like IBM are jumping on the LLM bandwagon and training/releasing models that have no chance of competing with Llama/Mistral? To me it just looks like a complete waste of $$ because nobody will use them in any serious scenarios

prob getting some reputation in AI space will help them to sell watsonx. tbf, watson predates Transformers paper.
Post reply on HN