alt.hn

8/18/2026 at 2:29:20 PM

Show HN: Shoehorn – Quantize any model down to run on your machine

https://notactuallytreyanastasio.github.io/shoehorn/

by rhgraysonii

8/18/2026 at 3:49:26 PM

Reminds me of https://github.com/AlexsJones/llmfit

by hmokiguess

8/18/2026 at 10:11:36 PM

LLMFit tells you what can run on something. I built something quite similar to their search into Shoehorn now.

by rhgraysonii

8/18/2026 at 3:19:34 PM

does this work similar to airllm? i am wondering how it would handle something like quantizing kimi k3 on a budget of 8 gbs, or is that something you are not attempting to solve yet?

by mbuchel-hn

8/18/2026 at 10:11:48 PM

Yes that is exactly what this does.

by rhgraysonii

8/19/2026 at 1:02:10 AM

Could you explain what happens when you try to shoehorn a 2.4T parameter model into a 24gb m4 mac?

by kennywinker

8/19/2026 at 2:09:42 PM

Wondering the same thing but for 48gb M5 Max.

by akshay_akula

8/18/2026 at 6:32:41 PM

tried it out but based on the model sizing result i got i got an insufficient memory error when the server started running

by jaylane

8/18/2026 at 10:12:06 PM

If you could post an issue if you still have the error around that would be awesome.

by rhgraysonii

8/20/2026 at 1:41:25 AM

[dead]

by kelvo_ran