Comment by cuttothechase
1 hour ago
Wondering how well this does with tool calling. Any one has any numbers or videos or anything using this?
From the github repo it seems like you really don't need a big Mac with huge amounts of RAM but SSD is sufficient.
If this is anywhere near 50 TPS, that would be a game changer in the personal LLM space!
No comments yet
Contribute on Hacker News ↗