← Back to context

Comment by HDBaseT

17 hours ago

In theory it sounds alright, although in reality, the heat and power consumption make this less than ideal.

I am not sure single digit tk/s on weak model like ChatGPT OSS 20b @ Q3_K_M is going to be worthwhile.

Not to mention this has to be <0.01% of users who "want" to do this. The iPhone tooling would make it hard to even configure and spin up a local LLM.