Comment by embedding-shape

3 hours ago

Hopper is ~4 years old at this point though, compared to Blackwell which is ~2 years old, the difference isn't nil.

Depending on your use case, you might prefer native FP4 and FP6 low-precision support and 5th-generation Tensor Cores rather than what Hopper offers.

I think for training Hopper makes sense as it's generally a bit cheaper and the difference isn't that big, but for inference the difference widens a bunch makes a lot more sense to go with Blackwell.

You're correct, the difference isn't nil - H100 is data-center grade GPU with loads of bandwidth and compute while RTX PRO is a consumer grade GPU. The difference between the microarchitectures would make sense to point out if we had been comparing apples to apples and not apples to pits.