Zebra-Llama – Towards efficient hybrid models
1–10 of 66 posts
Re: Zebra-Llama – Towards efficient hybrid models
#2Re: Zebra-Llama – Towards efficient hybrid models
#3This is an extraordinary claim, is there a catch I’m missing? Am I misreading?
Re: Zebra-Llama – Towards efficient hybrid models
#4that being said, I do very much believe that computational efficiency of models is going to go up [correction] drastically over the coming months, which does pose interesting questions over nvidia's throne
*previously miswrote and said computational efficiency will go down
Re: Zebra-Llama – Towards efficient hybrid models
#5Due to perverse incentives and the historical nature of models over-claiming accuracy, it's very hard to believe anything until it is open source and can be tested out that being said, I do very much believe that computational efficiency of models is going to go up [correction] drastically over the coming months, which does pose interesting questions over nvidia's throne *previously miswrote and said computational ef…
Re: Zebra-Llama – Towards efficient hybrid models
#6Due to perverse incentives and the historical nature of models over-claiming accuracy, it's very hard to believe anything until it is open source and can be tested out that being said, I do very much believe that computational efficiency of models is going to go up [correction] drastically over the coming months, which does pose interesting questions over nvidia's throne *previously miswrote and said computational ef…
Re: Zebra-Llama – Towards efficient hybrid models
#7Due to perverse incentives and the historical nature of models over-claiming accuracy, it's very hard to believe anything until it is open source and can be tested out that being said, I do very much believe that computational efficiency of models is going to go up [correction] drastically over the coming months, which does pose interesting questions over nvidia's throne *previously miswrote and said computational ef…
I think you mean computational efficiency will go _up_ in the future. To your last point: Jevons paradox might apply.
if computational efficiency goes up (thanks for the correction), and CPU inference becomes viable for most practical applications, GPUs (or accelerators) themselves may be unnecessary for most practical functions
Re: Zebra-Llama – Towards efficient hybrid models
#8Due to perverse incentives and the historical nature of models over-claiming accuracy, it's very hard to believe anything until it is open source and can be tested out that being said, I do very much believe that computational efficiency of models is going to go up [correction] drastically over the coming months, which does pose interesting questions over nvidia's throne *previously miswrote and said computational ef…
Like this? https://huggingface.co/amd/Zebra-Llama-8B-8MLA-24Mamba-SFT
Re: Zebra-Llama – Towards efficient hybrid models
#9Earlier quoted context omitted.
I think you mean computational efficiency will go _up_ in the future. To your last point: Jevons paradox might apply.
yup that's what I meant!, Jevon's paradox applies to resource usage in general and not towards a specific companies dominance if computational efficiency goes up (thanks for the correction), and CPU inference becomes viable for most practical applications, GPUs (or accelerators) themselves may be unnecessary for most practical functions