Comment by fwipsy
2 hours ago
Right, but datacenter GPUs optimized for LLM training/inference would have a bandwidth:compute ratio scaled to that workload.
2 hours ago
Right, but datacenter GPUs optimized for LLM training/inference would have a bandwidth:compute ratio scaled to that workload.
No comments yet
Contribute on Hacker News ↗