Comment by RandyOrion
7 days ago
Thanks Bonsai team.
Now open weight LLMs/VLMs/LMMs are becoming even larger to the extent that consumer-grade hardware are no longer able to run these models. In contrast, quantization and pruning make the model better at the size-performance pareto and provide people with strictly more possibilities.
No comments yet
Contribute on Hacker News ↗