Comment by wnmurphy

1 day ago

Yeah, I'm looking forward to this actually.

https://chatjimmy.ai/ blew my mind at how fast etched model weights can be.

For on-device LLMs, there's a point of diminishing returns, meaning you don't need to have the latest frontier model for most operations.

I currently have a small TTS model running in the background on my machine through which my agent(s) speak to me as they work. If that can be baked into an ASIC along with a few thousand voices in every major language then it should just be a utility chip on your mobo for anything that needs it. And yes, I too, am looking forward to it.