Live data from Hacker News

DeciLM-7B: The Fastest and Most Accurate 7B-Parameter LLM to Date

deci.ai

21–30 of 45 posts

Re: DeciLM-7B: The Fastest and Most Accurate 7B-Parameter LLM to Date

#22
Their website looks like buzzword bingo from the crypto days with no substance whitepapers and a big group of "researchers" + lots and lots of cool graphics and a huge money making "platform" before they have a product. "Book a demo" lmao.

The literal opposite of Mistral's "no marketing" just a torrent with the data.

Off course it may be actual new work but colour me very sceptical with this extreme confidence and marketingspeak.

Re: DeciLM-7B: The Fastest and Most Accurate 7B-Parameter LLM to Date

#23
post #17

Earlier quoted context omitted.

You have to look at the pre-trained models and not the "tuned" models. Its number one in that category (which is what I think they are referring to, given the benchmarks are against Mistral and llama)

Doesn't the article say this is a LoRA off ORCA? Isn't that a finetune?

DeciLM-7B-instruct is the finetuned model, not DeciLM-7B.

Re: DeciLM-7B: The Fastest and Most Accurate 7B-Parameter LLM to Date

#24

Their website looks like buzzword bingo from the crypto days with no substance whitepapers and a big group of "researchers" + lots and lots of cool graphics and a huge money making "platform" before they have a product. "Book a demo" lmao. The literal opposite of Mistral's "no marketing" just a torrent with the data. Off course it may be actual new work but colour me very sceptical with this extreme confidence and ma…

Sure, but a crucial difference here is you can go right now and try it out and see if it's garbage or not on HuggingFace[0]

0: https://huggingface.co/spaces/Deci/DeciLM-7B-instruct

Re: DeciLM-7B: The Fastest and Most Accurate 7B-Parameter LLM to Date

#25

The language in this blog post is overblown to say the least. "groundbreaking", "outshines its competitors", "remarkable", "pivotal transformation" Totally unnecessary bravado.

Why is it that so many people here have to comment on basic marketing stuff? It is necesssary bravado because it's trying to convince people that it's worth using. And saying "Here is a model that is just a bit better than others" isn't going to do anything. Therefore it's necessary to have buzzwords such as groundbreaking.

Re: DeciLM-7B: The Fastest and Most Accurate 7B-Parameter LLM to Date

#27
post #21

The language in this blog post is overblown to say the least. "groundbreaking", "outshines its competitors", "remarkable", "pivotal transformation" Totally unnecessary bravado.

Maybe it is written by the model itself.

Hah, the model is pretty good, but the material does sound like an American politician’s hype team. To be fair I think google and OAI are doing the same thing all the time:

“As a large language model, I can explain. I and my relative were trained at the bestest universities, only in superlatives about humans who are optimum et vetustissimum octagenarium, i.e. the “bestest of all time! Could win nathan’s hot dog-eating contest while making the Jersey Turnpike Marathon unfair and telling you about how it used to cost a nickel at the ferry. Everyone else stinks stonks!””

Re: DeciLM-7B: The Fastest and Most Accurate 7B-Parameter LLM to Date

#28
Is anyone actually using these small models for general purpose applications? Or are they just used to fine tune into narrowly useful specialized models?

I keep seeing them pass each other on different benchmarks and leaderboards and I can't help but imagine that so many of them are only good at benchmarks and not much else.

I haven't gotten a chance to play with this stuff yet so I have no basis to go off of, but I'd like to hear from anyone who's actually been impressed by any of these small models for general purpose tasks.

Re: DeciLM-7B: The Fastest and Most Accurate 7B-Parameter LLM to Date

#29

Ignoring the fact that this is a 1 day old account, why should I care that this LLM's bars are larger and numbers bigger? Is this corporate blogpost actually good content for anyone? What is the appeal other than "yup, that blue bar sure is longer than the other color bars."?

It's fine if you don't care. Nobody needs to care about everything. But why post your comment? It doesn't add anything to discussion. For me I care about different companies reaching mistral level so that mistral or whoever is in top has to release the model weights else competitors will.

I posted my comment because I see a new one of these on the front page every morning and I was genuinely curious if there was anyone who cares about these posts? That's alongside the fact I'm almost certain that the account that posted this is a sockpuppet for the company advertising this LLM.

Re: DeciLM-7B: The Fastest and Most Accurate 7B-Parameter LLM to Date

#30

Is anyone actually using these small models for general purpose applications? Or are they just used to fine tune into narrowly useful specialized models? I keep seeing them pass each other on different benchmarks and leaderboards and I can't help but imagine that so many of them are only good at benchmarks and not much else. I haven't gotten a chance to play with this stuff yet so I have no basis to go off of, but I'…

Mistral-7B is surprisingly decent for general purpose small tasks. The more complex the task or the more specific the knowledge recall, the worse the performance since the smaller the models are - the less breadth they tend to have.

But they're very nice for making PoCs on complex systems since they're near free to run.

Post reply on HN