Live data from Hacker News

StableLM Zephyr 3B

stability.ai

41–42 of 42 posts

Re: StableLM Zephyr 3B

#41
post #16

Am I reading it right that performance was roughly comparable with GPT-3.5? How is this even possible?

The LLM field is still messy at large, if you look at the rankings of model performance, they still do not reflect their usability in real life. I think one major challenge is to find a corresponding benchmark.

Re: StableLM Zephyr 3B

#42
post #35
post #26

Earlier quoted context omitted.

It’ll be static with cpi max, flat membership for all core models Just released video, sdxl turbo and 3d, code and more coming Very positive reaction so far and we will still do our grants for OSS and do OSS collaborations, done over 10m A100 hours over last year Launches next few days https://x.com/emostaque/status/1732197072290353455?s=46

You forgot the part where a holding company eventually buys Stability and starts to squeeze their customers that built the core part of their product around a pricing structure that's "static + CPI". With a bit less snark; I think the pricing structure is reasonable, but that doesn't mean customers won't eventually get screwed over by it. See the Unity fiasco from earlier this year (which was backed out, but serves w…

> See the Unity fiasco from earlier this year (which was backed out, but serves well as an example).

They learned from the Russians. The Cuban missile crisis wasn't about Russian missiles in Cuba, it was about establishing an air base in Cuba. The missiles were the "price shock" item, and the airbase was the goal. If the Russians had put in just an airbase, it was very likely they never would have kept it.

Post reply on HN