Comment by ixaxaar 2 days ago I'm running it using a 4090 on using llama.cpp with Q5_K_S and its running at ~33 t/s 0 comments ixaxaar Reply No comments yet Contribute on Hacker News ↗
No comments yet
Contribute on Hacker News ↗