Comment by noosphr 5 hours ago Issue is that llama.cpp is the best way to run models on hardware that isn't nvidias. 1 comment noosphr Reply alightsoul 5 hours ago Except when they have less than 16 gb of ram?
Except when they have less than 16 gb of ram?