Live data from Hacker News

SubQ 1.1 Small

subq.ai

51–54 of 54 posts

Re: SubQ 1.1 Small

#51

This one's interesting, and I think the next frontier for LLMs should really just be, how can we get something like Opus 4.6 to cost drastically less, for the same output? I say 4.6 because from 4.6 onwards it's been pretty darn good, at least for me, always feels like every model upgrade someone hates it, heck even 4.5 was fine.

I agree on opus 4.5-4.8, but Fable 5 was a noticeable upgrade.

Man I miss it

Re: SubQ 1.1 Small

#52

Earlier quoted context omitted.

It's not quite true to say that if you release it you get nothing. If it's worthwhile and picked up by the open-weights labs, you get much bigger and better models implementing it than you would have had access to or been able to train otherwise, quicker than if they had to figure it out de novo.

Yeah. I am about to the point of just releasing it all. I love the tech. It does amazing things. But I want to move to the next big things I can see doing with it and building the custom ops to get it to work efficiently is a pain. I am positive others would run with it and make it all way better which would free me up to do more.

well if you ever release it, make sure to make a post so we can check it out !

Re: SubQ 1.1 Small

#53
"At 1M tokens, SubQ 1.1 Small requires 64.5x less compute than dense attention and runs 56x faster than FlashAttention-2."

6450% less compute? Is Trump working there?

Post reply on HN