Live data from Hacker News

Show HN: Shoehorn – Quantize any model down to run on your machine

notactuallytreyanastasio.github.io

1–10 of 21 posts

Show HN: Shoehorn – Quantize any model down to run on your machine

#1
Working on Mac, Linux, and Windows now. I include a simple GUI to find new models and get things built and set up. It is working quite well across a few models for me. The GitHub README and DESIGN.md files go into detail of the how/why and it's working remarkably well so far. https://github.com/notactuallytreyanastasio/shoehorn

Show HN: Shoehorn – Quantize any model down to run on your machine
notactuallytreyanastasio.github.io

Re: Show HN: Shoehorn – Quantize any model down to run on your machine

#6

does this work similar to airllm? i am wondering how it would handle something like quantizing kimi k3 on a budget of 8 gbs, or is that something you are not attempting to solve yet?

Yes that is exactly what this does.

Re: Show HN: Shoehorn – Quantize any model down to run on your machine

#8

does this work similar to airllm? i am wondering how it would handle something like quantizing kimi k3 on a budget of 8 gbs, or is that something you are not attempting to solve yet?

Yes that is exactly what this does.

Could you explain what happens when you try to shoehorn a 2.4T parameter model into a 24gb m4 mac?

Re: Show HN: Shoehorn – Quantize any model down to run on your machine

#10

Earlier quoted context omitted.

Yes that is exactly what this does.

Could you explain what happens when you try to shoehorn a 2.4T parameter model into a 24gb m4 mac?

Wondering the same thing but for 48gb M5 Max.
Post reply on HN