Comment by janalsncm

2 months ago

For the most part it’s better than Nemotron, worse than GLM. This makes it the best American open weights model from what I can tell?

It's nearly double the size of Nemotron 3 Ultra, so I'd expect it to be considerably better, although the active parameter count seems to be a touch lower at 41B vs 55B

I'm surprised that Nemotron gets mentioned at all. In my experiments with it for coding tasks it performed extremely poorly, essentially unusable.

  • I focus on realtime voice AI uses cases and nemotron's time to first token is INSANELY fast. It's become a legit option for voice use cases