I wonder why companies like IBM are jumping on the LLM bandwagon and training/releasing models that have no chance of competing with Llama/Mistral? To me it just looks like a complete waste of $$ because nobody will use them in any serious scenarios
I've not seen any proper evaluations for Granite against, say, Llama or Mistral.
Until we do it's probably too early to say they can't compete, at least in some areas where others perform poorly.