Live data from Hacker News

Show HN: Shoehorn – Quantize any model down to run on your machine

notactuallytreyanastasio.github.io

11–20 of 21 posts

Re: Show HN: Shoehorn – Quantize any model down to run on your machine

#14

I gotta laugh at some of the models it suggests, for example: > AnkitAI/Parable-Qwen3-4B-Claude-Fable-5-GGUF you’re telling me you managed to fit Fable 5 into just 4B?

[flagged]

Don't make fun of people you think are ignorant, it's a pretty shitty look

Re: Show HN: Shoehorn – Quantize any model down to run on your machine

#17

I gotta laugh at some of the models it suggests, for example: > AnkitAI/Parable-Qwen3-4B-Claude-Fable-5-GGUF you’re telling me you managed to fit Fable 5 into just 4B?

[flagged]

Well to be fair here... the title of this post doesn't mention fine tuning, it mentions quantization.

Re: Show HN: Shoehorn – Quantize any model down to run on your machine

#18

Earlier quoted context omitted.

Yes that is exactly what this does.

Could you explain what happens when you try to shoehorn a 2.4T parameter model into a 24gb m4 mac?

extreme divergence would be my guess

Re: Show HN: Shoehorn – Quantize any model down to run on your machine

#20

I gotta laugh at some of the models it suggests, for example: > AnkitAI/Parable-Qwen3-4B-Claude-Fable-5-GGUF you’re telling me you managed to fit Fable 5 into just 4B?

Fyi that model name to me reads

Qwen3 4b params distilled/trained with fable 5

Post reply on HN