← Back to context

Comment by itvision

10 hours ago

OMFG this thing is fast.

its fast but try to get it to give you pi to 50 decimal places. it didnt go well for me.

  • I think the same exact model running on CPU-only and RAM, or a small GPU, would do about the same? It's quite an old model now and small, you could throw a GGUF into llama-server or something for a side by side comparison.

    https://huggingface.co/meta-llama/Llama-3.1-8B

    As I remember just about any english language model from mid 2024 and earlier didn't even do well if you asked it to count sequentially from 0 to 100, nevermind calculating stuff.