NHacker Next
- new
- past
- show
- ask
- show
- jobs
- submit
login
Reminds me of https://github.com/AlexsJones/llmfit
LLMFit tells you what can run on something. I built something quite similar to their search into Shoehorn now.
This is interesting. I wonder how it could work with something like https://github.com/JustVugg/colibri.
does this work similar to airllm? i am wondering how it would handle something like quantizing kimi k3 on a budget of 8 gbs, or is that something you are not attempting to solve yet?
Yes that is exactly what this does.
Could you explain what happens when you try to shoehorn a 2.4T parameter model into a 24gb m4 mac?
Wondering the same thing but for 48gb M5 Max.
tried it out but based on the model sizing result i got i got an insufficient memory error when the server started running
If you could post an issue if you still have the error around that would be awesome.
[dead]
Rendered at 09:44:44 GMT+0000 (Coordinated Universal Time) with Vercel.