Comment by adrian_b
2 months ago
There is also NVIDIA Nemotron 3 Ultra, with 561B parameters, which was released a month ago.
I do not know yet how smart it is, but the NVIDIA LLMs are very well optimized for fast inference (on their GPUs of course).
Previously that was the biggest American open-weights LLM.
No comments yet
Contribute on Hacker News ↗