NHacker Next
  • new
  • past
  • show
  • ask
  • show
  • jobs
  • submit
Show HN: Shoehorn – Quantize any model down to run on your machine (notactuallytreyanastasio.github.io)
hmokiguess 2 days ago [-]
rhgraysonii 1 days ago [-]
LLMFit tells you what can run on something. I built something quite similar to their search into Shoehorn now.
akshay_akula 20 hours ago [-]
This is interesting. I wonder how it could work with something like https://github.com/JustVugg/colibri.
mbuchel-hn 2 days ago [-]
does this work similar to airllm? i am wondering how it would handle something like quantizing kimi k3 on a budget of 8 gbs, or is that something you are not attempting to solve yet?
rhgraysonii 1 days ago [-]
Yes that is exactly what this does.
kennywinker 1 days ago [-]
Could you explain what happens when you try to shoehorn a 2.4T parameter model into a 24gb m4 mac?
akshay_akula 20 hours ago [-]
Wondering the same thing but for 48gb M5 Max.
jaylane 2 days ago [-]
tried it out but based on the model sizing result i got i got an insufficient memory error when the server started running
rhgraysonii 1 days ago [-]
If you could post an issue if you still have the error around that would be awesome.
kelvo_ran 8 hours ago [-]
[dead]
Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact
Rendered at 09:44:44 GMT+0000 (Coordinated Universal Time) with Vercel.